influxdb

Commit Graph

Author	SHA1	Message	Date
Jonathan A. Sternberg	c8c38e15cd	Merge pull request #6386 from influxdata/js-iterator-next-error Modify all of the iterators to allow returning an error on Next()	2016-04-20 10:39:53 -04:00
Ben Johnson	54454e1e5b	Merge pull request #6424 from benbjohnson/optimize-bit-reader Optimize tsm1.BitReader	2016-04-20 08:28:24 -06:00
Seif Lotfy	c6e3c87e00	Add Block checksum validation and "influx_inspect verify" tool Fixes #5502	2016-04-19 22:33:03 +02:00
Ben Johnson	1d2238c642	optimize tsm1.BitReader This commit rewrites the `tsm1.BitReader` to use an 8-byte buffer instead of a 1-byte buffer and provide an inlineable fast bit read.	2016-04-19 11:34:17 -06:00
Jason Wilder	f841a90d35	Use int64 instead of time.Time in timestamp encoder/decoder	2016-04-19 10:25:27 -06:00
Jason Wilder	61beeca426	Update timestamp benchmarks	2016-04-19 10:17:32 -06:00
Jonathan A. Sternberg	7ec2a991d5	Modify all of the iterators to allow returning an error on Next() This also switches the remaining iterators to be lazy so they can return errors properly. They needed to be converted to lazy initialization anyway, which has the side effect of making it much easier for us to propagate the underlying error during initialization. Updated the Emitter to return errors when it cannot read properly from the iterators.	2016-04-18 11:17:55 -04:00
Jonathan A. Sternberg	93745d9693	Merge pull request #6391 from influxdata/js-5553-limit-queries-slow-with-group-by Propagate the limit option to the low level iterators	2016-04-16 09:39:25 -04:00
Jonathan A. Sternberg	bd5fdd797d	Propagate the limit option to the low level iterators When a GROUP BY or multiple sources are used, the top level limit iterator requires reading the entire iterator stream so it can find all of the tag groups it needs to return. For large data series, this ends up with the limit iterator discarding a lot of output. This change adds a new lower level limit iterator on each series itself so that there are fewer data points that have to be thrown away by the top level iterator. Fixes #5553.	2016-04-15 18:23:54 -04:00
Jonathan A. Sternberg	835d08591e	Do not filter out empty tags from series keys	2016-04-13 09:15:57 -04:00
Jonathan A. Sternberg	60282cf52d	Merge pull request #6284 from influxdata/js-3371-where-clause-compare-tags-and-fields Enhance comparing tags and fields in the where clause	2016-04-12 11:45:54 -04:00
Pierre Fersing	29b19a2293	Fix deadlock in tsm1/file_store	2016-04-12 09:39:21 +02:00
Jonathan A. Sternberg	ea6262b712	Enhance comparing tags and fields in the where clause Now it is possible to compare tags and fields and it is also now possible to compare tags and tags. Previously, it was only possible to compare fields with fields and tags with a string or a regex. Fixes #3371.	2016-04-11 18:10:08 -04:00
Ben Johnson	525e22c92b	tsm1 query engine alloc reduction This commit makes a number of performance improvements to reduce allocations during query execution. Several objects and buffers are now reused across the components to avoid allocations. Previously a simple `count(value)` query across 1M points would require 26,000+ allocations. After the changes in this commit that number has been reduced to 88.	2016-04-11 14:50:59 -06:00
Jonathan A. Sternberg	028fdaff81	Merge pull request #6222 from influxdata/js-6206-descending-tsm1-iterators Handle nil values from the tsm1 cursor correctly	2016-04-06 10:05:20 -04:00
Jonathan A. Sternberg	94ec92d669	Handle nil values from the tsm1 cursor correctly Send nil values from the tsm1 cursor at the end of the cursor. After the cursor reached tsm1, the `nextAt()` call would always return the default value rather than a nil value. Descending also didn't work correctly because the seeking functionality for tsm1 iterators would always act like they were ascending instead of descending when choosing which value to select. This resulted in very strange output from the emitter since it couldn't figure out if it was ascending or descending. Fixes #6206.	2016-04-06 09:27:02 -04:00
Jason Wilder	3f4c5a5585	Fix race on measurementFields Both Shard and Engine had the same reference to the measurementField map, but they each protected it with their own locks. This causes a race when write and queries are occurring because writes can add new fields to the map while queries are reading from it. The fix moves the ownership to the Engine and provides protected accessors to that Shard now users. For the most parts, the access on shard were old dead code. Fixing the measurementFields map race created a new race on the internal fields map. This is now unexported and protected via MeasurementFields exported funcs. Fixes #6188	2016-04-01 18:57:01 -06:00
Jason Wilder	873ac2715d	Fix panic: runtime error: slice bounds out of range Writing a key that exceeds the max key length could cause a panic when reading a tsm file because the 2 bytes used for the key length would not be enough to represent the actual key length. The writer will now return an error if when trying to write a key that is too large.	2016-03-30 23:44:17 -06:00
Jonathan A. Sternberg	711a6614e6	Implement the point limit monitor Fixes #6077.	2016-03-30 16:08:56 -04:00
Joe LeGasse	f10c300765	Update to conversion tool to work in current versions After adding type-switches to the tsm1 packages, the custom implementation found in the conversion tool broke. This change uses tsm1.NewValue() instead of a custom implementation. This change also ensures that the tsm1.Value interface can only be implemented internally to allow for the optimized type-switch based encoding	2016-03-30 13:26:46 -04:00
Jason Wilder	60c3898577	Add godoc for KeyAt func	2016-03-29 12:59:26 -06:00
Jason Wilder	1b08e2dd55	Use walk func to load all tsm keys to index Avoids allocating a big map or all keys.	2016-03-29 12:59:26 -06:00
Jason Wilder	d4757ad040	Remove sync.Pool from wal UnmarshalBinary When loading many shards concurrently they block trying to acquire a write lock in the sync pool adding a new source of contention. Since this code flow always needs to allocate a buffer it's not really buying us much.	2016-03-29 12:59:26 -06:00
Jason Wilder	03ced4cc90	Load shards concurrently	2016-03-29 12:58:52 -06:00
Ben Johnson	45f1c28adb	add tsm iterator stats buffer This commit adds a buffer for stats to be updated without requiring a mutex lock/unlock on every point. The tradeoff is that stats are not exactly precise. This works for our use case because stats are only periodically checked.	2016-03-23 12:23:22 -06:00
Jonathan A. Sternberg	a35d9602cd	Fix where filters when a OR is used and when a tag does not exist If an OR was used, merging filters between different expressions would not work correctly. If one of the sides had a set of series ids with a condition and the other side had no series ids associated with the expression, all of the series from the side with a condition would have the condition ignored. Instead of defaulting a non-existant series filter to true, it should just be false and the evaluation of the one side that does exist should take care of determining if the series id should be included or not. The AND condition used false correctly so did not have to be changed. If a tag did not exist and `!=` or `!~` were used, it would return false even though the neither a field or a tag equaled those values. This has now been modified to correctly return the correct series ids and the correct condition. Also fixed a panic that would occur when a tag caused a field access to become unnecessary. The filter using the field access still got created and used even though it was unnecessary, resulting in an attempted access to a non-initialized map. Fixes #5152 and a bunch of other miscellaneous issues.	2016-03-22 12:19:06 -04:00
Ben Johnson	6e1c1da25b	reduce allocations in query execution This commit removes some heap objects by converting them from pointer references to non-pointers or by reusing buffers.	2016-03-22 09:51:39 -06:00
Jonathan A. Sternberg	ad96207868	Fix ORDER BY desc so it doesn't skip values After reading the initial buffer, ORDER BY desc would read the next block into the buffer and only read the first element. It's because the code that was copied from the ascending cursor wasn't modified correctly to set the position to the last element in the buffer. The buffer size has also been lowered from 1000 to 10 to match with the ascending cursor for performance with limit queries. Fixes #6055.	2016-03-22 09:40:11 -04:00
Ben Johnson	7156c1f9bd	add IteratorStats This commit adds an `IteratorStats` that holds aggregate iterator processing information. A method is also added to `Iterator` to return the stats: Stats() influxql.IteratorStats The remote iterators will also emit their stats in the point stream upon first connection, on a given interval, and then finally once the last point has been sent.	2016-03-21 16:25:19 -06:00
Jason Wilder	ee2f21e76f	Merge pull request #6082 from influxdata/jw-tsm Fix partially written TSM files	2016-03-21 15:42:27 -06:00
Jason Wilder	7567453c9a	Ensure TSM files are fsync'd Make sure TSM files are fsync'd when closed and also that the parent dir is fsync'd when they are renamed.	2016-03-21 15:03:52 -06:00
Jason Wilder	a4e5446ddd	Return error when TSM writer close returns one The TSM writer uses a bufio.Writer that needs to be flushed before it's closed. If the flush fails for some reason, the error is not handled by the defer and the compactor continues on as if all is good. This can create files with truncated indexes or zero-length TSM files. Fixes #5889	2016-03-21 15:00:36 -06:00
Jonathan A. Sternberg	6655ca7769	Create a new interrupt iterator that will stop emitting points after an interrupt Use of the iterator is spread out into both `IteratorCreators` and inside of the iterators themselves. Part of the interrupt must be handled inside of the engine so it stops trying to emit points when an interrupt is found and another part of the interrupt has to happen when combining the iterators so it doesn't just start reading the next shard.	2016-03-21 12:07:07 -04:00
Jason Wilder	3fd40d48a1	Merge pull request #6006 from influxdata/jw-deadlock Fix deadlock when running backup	2016-03-14 13:36:45 -06:00
Jason Wilder	9984cd5d6d	Fix skipping blocks at query time when overlaps exist Depending on how data is written across TSM files, it was possible to skip over some blocks at query time making it looks like data was missing.	2016-03-14 13:11:11 -06:00
Jason Wilder	000459e350	Fix deadlock when running backup A deadlock occurs under write load if a backup is run in between the time when a snapshot compactions has snapshotted the cache and successfully written it to disk. The issus is that the second snapshot call will block on the commit lock while it is holding the engine write lock. This causes all writes to block as well as prevents the currently runnign snapshot compaction from completing because it needs to acquire a read-lock. This PR removes the commit lock and just returns an error if a snapshot is in progress to all any locks being held to be released. The caller can determine whether to retry or giveup.	2016-03-14 12:36:48 -06:00
Joe LeGasse	344e5abd41	Changed type-switch a few places to reduce allocations. Slices of tsm1.Value interfaces are only ever used with all the same types, and the previous code would switch on the type returned from a call to Value(), which allocated and returned an interface{} object for the underlying value. This change instead type-switches on the tsm1.Value object itself, allowing it direct access to the underlying value field, eliminating the unecessary allocations.	2016-03-11 15:57:05 -05:00
Jason Wilder	992c78ee22	Remove period shard maintenance goroutine This is no longer used in tsm and just peridocially locks everything for no reason now.	2016-03-09 17:31:02 -07:00
Edd Robinson	58c03448aa	Merge pull request #5514 from influxdata/er-engine-panic Ensure shards and engine are safely closed	2016-03-09 18:56:36 +00:00
Jason Wilder	e3fef5593c	Merge pull request #5855 from jonseymour/jss-5854-go-master-breaks-build fix tests to cope with future changes to testing.quick.Check - see #5854	2016-03-01 19:03:21 -07:00
Mark Rushakoff	cdcb079769	Tag TSM stats with database, retention policy ... by extracting the db/rp from the given path. Now that the code has "standardized" on extracting db/rp this way, the ShardLocation struct is no longer necessary and thus has been removed. We're back on the previous style of passing the path and walPath to NewShard.	2016-02-29 09:17:34 -08:00
Jon Seymour	73b3a2a056	Merge #5855 (issue: #5854 ). RHS merges cleanly with 0.10.0 Signed-off-by: Jon Seymour <jon@wildducktheories.com>	2016-02-29 20:37:32 +11:00
Jon Seymour	716cdd7f41	tsm: modify encoding tests to deal with possible nil slices from testing.quick.Check in go master The current go compiler at the tip of the go master (1d5001af) has a modified implementation of testing.quick.Check that now generates nil slices as test data. (See: https://gophers.slack.com/archives/general/p14567053570110). The existing tests expect round tripping to work in this case but it does not. So, in these cases we change the expectation to reflect actual behaviour. This needs to be checked for reasonableness.	2016-02-29 20:36:19 +11:00
Jason Wilder	8d70d65a82	Convert time.Time to int64	2016-02-25 15:15:01 -07:00
Jon Seymour	11123d2694	Merge #5833 (issue: #5832 ). Signed-off-by: Jon Seymour <jon@wildducktheories.com>	2016-02-26 07:59:03 +11:00
Jon Seymour	2c7cd06b99	tsm: cache: need to check that snapshot has been sorted. Previously, the for loop at the end of the method assumed that all entries had been deduplicated, including the entry discovered in the snapshot. However, this wasn't actually true. With this change, we make it true. Signed-off-by: Jon Seymour <jon@wildducktheories.com>	2016-02-26 07:56:25 +11:00
Jon Seymour	7eabae68de	tsm: cache: add a test for the write sequence {6,1,snapshot,7,2} Consider the write sequence: 6,1,snapshot,7,2. The hot cache gets deduplicated, so is 2,7. Now consider the test if 1 >= 2, this is false, so needSort is not set to true. The problem is the implicit assumption that the snapshot is always sorted by the time that merged() runs, but this may not be true. Signed-off-by: Jon Seymour <jon@wildducktheories.com>	2016-02-26 07:43:50 +11:00
Jason Wilder	6ebc192298	Merge pull request #5678 from jonseymour/typo doc: typographical, spelling, grammar, word-choice and phrasing improvements.	2016-02-25 09:33:41 -07:00
Jason Wilder	daf68dbbd2	Merge pull request #5701 from jonseymour/js-deduplicate-safety tsm: cache: improve thread safety of Cache.Deduplicate (see #5699)	2016-02-25 09:18:10 -07:00
Jon Seymour	4d98a1cf28	tsm: cache: remove unnecessary lock escalation. Previously, we needed a write lock on the cache because it was the only lock we had available to guard updates to entry.values and entry.needSort. However, now we have a entry-scoped lock for this purpose, we don't need the cache write lock for this purpose. Since merged() doesn't modify the .store or the c.snapshot.sort, there is no need for a write lock on the cache to protect the cache. So, we don't need to escalate here - we simply rely on the entry lock to protect the entries we are iterating over. Signed-off-by: Jon Seymour <jon@wildducktheories.com>	2016-02-26 01:31:54 +11:00

1 2 3 4 5 ...

520 Commits (e3275f22b730a06b6cd958fa55822a0d7f075a6d)