influxdb

Commit Graph

Author	SHA1	Message	Date
Cory LaNou	0103e44896	allow partial writes on field conflicts	2017-01-23 12:25:35 -06:00
Ben Johnson	c459d24a60	Test coverage.	2017-01-23 09:38:27 -07:00
Gunnar	3722fa383d	Merge pull request #7718 from influxdata/ga-drop-stats Add stats on dropped measurements and series; Fixes #7697	2017-01-20 15:54:06 -08:00
Edd Robinson	feb7a2842c	Use unbuffered error channels in tests	2017-01-17 10:53:15 -08:00
Edd Robinson	fb7388cdfc	Remove dead code from various pkgs	2017-01-17 09:47:34 -08:00
Edd Robinson	292b30b82b	Fix subtle bugs and remove dead code from tsdb	2017-01-17 09:47:34 -08:00
Edd Robinson	320c5981cb	Fixes racy locking on measurement	2017-01-17 09:44:56 -08:00
Edd Robinson	45324b3848	Fixes racy locking on measurement	2017-01-16 14:22:11 -08:00
Joe LeGasse	cd00085e9e	Adjust Tags cloning This change delays Tag cloning until a new series is found, and will only clone Tags acquired from `ParsePoints...` and not those referencing the mmap-ed files (TSM) that are created on startup.	2017-01-13 13:15:36 -05:00
Mark Rushakoff	cdbdd156f3	Fix memory leak of retained HTTP write payloads This leak seems to have been introduced in `8aa224b22d`, present in 1.1.0 and 1.1.1. When points were parsed from HTTP payloads, their tags and fields referred to subslices of the request body; if any tag set introduced a new series, then those tags then were stored in the in-memory series index objects, preventing the HTTP body from being garbage collected. If there were no new series in the payload, then the request body would be garbage collected as usual. Now, we clone the tags before we store them in the index. This is an imperfect fix because the Point still holds references to the original tags, and the Point's field iterator also refers to the payload buffer. However, the current write code path does not retain references to the Point or its fields; and this change will likely be obsoleted when TSI is introduced. This change likely fixes #7827, #7810, #7778, and perhaps others.	2017-01-12 16:16:54 -08:00
Joe LeGasse	2db0250b22	Add db/rp name validation This change adds some very basic name validation with the following plain-english description: names must be non-zero sequence of printable characters that do not contain slashes ('/' or '\') and are not equal to either "." or "..". The intent is that, since we currently just use database and retention policy names directly as path elements, these rules will hopefully leave us with names that should be at least close to valid directory names. Ideally, we would restrict names even further or not use them as path elements directly, but this should be a step towards the former without restricting names "too much"	2017-01-12 17:38:10 -05:00
Joe LeGasse	b19260fb26	Add some checks before removing directories Fixes #7822 This change first ensures that databases and retention policies exist before attempting to remove them from the Store. It also adds some checks in the `DeleteDatabase` and `DeleteRetentionPolicy` to ensure that maliciously named entries won't remove anything outside of the configured data directory.	2017-01-12 17:38:10 -05:00
Joe LeGasse	bf58d9ffb7	Update backup to use ioutil.ReadDir	2017-01-12 16:28:01 -05:00
Jason Wilder	11f264563a	Fix 32bit alignment	2017-01-12 12:01:49 -07:00
Jason Wilder	06a8fd6ca2	Simplifications and cleanup	2017-01-12 09:55:38 -07:00
Ben Johnson	f43b0f7ec9	Fix series & measurement deletion.	2017-01-12 09:29:40 -07:00
Edd Robinson	73ed864e1d	Add cache tests	2017-01-12 16:27:16 +00:00
Jason Wilder	1e56b5416b	Fix compactions sometimes getting stuck I ran into an issue where the cache snapshotting seemed to stop completely causing the cache to fill up and never recover. I believe this is due to the the Timer being reused incorrectly. Instead, use a Ticker that will fire more regularly and not require the resetting logic (which was wrong).	2017-01-11 17:57:40 -07:00
Jason Wilder	40b017f4a4	Fix Cache stats size collection The memory stats as well as the size of the cache were not accurate. There was also a problem where the cache size would be increased optimisitically, but if the cache size limit was hit, it would not be decreased. This would cause the cache size to grow without bounds with every failed write.	2017-01-11 17:54:51 -07:00
Jason Wilder	c433ff331f	Encode snapshots concurrently The CacheKeyIterator (used for snapshot compactions), iterated over each key and serially encoded the values for that key as the TSM file is written. With many series, this can be slow and will only use 1 CPU core even if more are available. This changes it so that the key space is split amongst a number of goroutines that start encoding all keys in parallel to improve throughput.	2017-01-11 17:54:27 -07:00
Jason Wilder	ae838ef323	Simplify Cache.Snapshot This simplifies the cache.Snapshot func to swap the hot cache to the snapshot cache instead of copy and appending entries. This reduces the amount of time the cache is write locked which should reduce cache contention for the read only code paths.	2017-01-11 11:12:02 -07:00
Jonathan A. Sternberg	3ba950b029	Fix for subqueries to use the parallel iterator correctly Also, fix the `Iterators.Merge(IteratorOptions)` function so it consults the `Ordered` attribute to determine which iterator it should use to merge the input iterators.	2017-01-11 10:47:18 -06:00
Ben Johnson	352817e8c4	Convert 32-bit offsets to 64-bit.	2017-01-11 08:58:10 -07:00
Jonathan A. Sternberg	b58d1778e2	Remove improper newlines from logging statements	2017-01-10 11:20:09 -06:00
Mark Rushakoff	a135906b43	Merge pull request #7747 from influxdata/mr-lint-cleanup Miscellaneous lint cleanup	2017-01-10 08:22:00 -08:00
Mark Rushakoff	3b3604e362	Fix race in (*tsm1.Cache).values Without this read lock, this race would happen during a concurrent snapshot compaction and query.	2017-01-09 14:48:28 -08:00
Jonathan A. Sternberg	4a559c4620	Merge pull request #7646 from influxdata/js-4619-subqueries Support subquery execution in the query language	2017-01-09 14:14:01 -06:00
Jason Wilder	eb4d311c0a	Add retry/backup when backing up a shard fails The backup command can fail if a snapshot is running which silently closes the connection. This causes the backup shard command to continue on as if nothing failed.	2017-01-09 11:28:48 -07:00
Ben Johnson	64c7715243	Rebase fixes.	2017-01-09 10:10:12 -07:00
Jason Wilder	194c5adfaf	Fix race on t.refs Read at 0x00c42018f620 by goroutine 58: github.com/influxdata/influxdb/tsdb/engine/tsm1.(TSMReader).Close() /root/go/src/github.com/influxdata/influxdb/tsdb/engine/tsm1/reader.go:330 +0x94 github.com/influxdata/influxdb/tsdb/engine/tsm1.(FileStore).Close() /root/go/src/github.com/influxdata/influxdb/tsdb/engine/tsm1/file_store.go:464 +0x123 Previous write at 0x00c42018f620 by goroutine 63: sync/atomic.AddInt64() /usr/local/go/src/runtime/race_amd64.s:276 +0xb github.com/influxdata/influxdb/tsdb/engine/tsm1.(TSMReader).Unref() /root/go/src/github.com/influxdata/influxdb/tsdb/engine/tsm1/reader.go:352 +0x43 github.com/influxdata/influxdb/tsdb/engine/tsm1.(KeyCursor).Close()	2017-01-07 12:39:45 -07:00
Jonathan A. Sternberg	d7c8c7ca4f	Support subquery execution in the query language This adds query syntax support for subqueries and adds support to the query engine to execute queries on subqueries. Subqueries act as a source for another query. It is the equivalent of writing the results of a query to a temporary database, executing a query on that temporary database, and then deleting the database (except this is all performed in-memory). The syntax is like this: SELECT sum(derivative) FROM (SELECT derivative(mean(value)) FROM cpu GROUP BY *) This will execute derivative and then sum the result of those derivatives. Another example: SELECT max(min) FROM (SELECT min(value) FROM cpu GROUP BY host) This would let you find the maximum minimum value of each host. There is complete freedom to mix subqueries with auxiliary fields. The only caveat is that the following two queries: SELECT mean(value) FROM cpu SELECT mean(value) FROM (SELECT value FROM cpu) Have different performance characteristics. The first will calculate `mean(value)` at the shard level and will be faster, especially when it comes to clustered setups. The second will process the mean at the top level and will not include that optimization.	2017-01-07 13:00:48 -06:00
Mark Rushakoff	153277c01d	Merge pull request #7786 from influxdata/mr-cache-decrease-size Use one atomic operation in (*Cache).decreaseSize	2017-01-06 10:17:01 -08:00
Ben Johnson	2b3cd415e2	Fixing rebase.	2017-01-06 09:52:16 -07:00
Ben Johnson	d1f1e19591	Fixing rebase.	2017-01-06 09:31:25 -07:00
Ben Johnson	1003db0067	Add active log file tracking, time-based compaction.	2017-01-05 10:17:12 -07:00
Ben Johnson	c1c98223ec	Fix and optimize tsi1 FileSet.	2017-01-05 10:17:12 -07:00
Ben Johnson	31e74d809b	Add tsi FileSet.	2017-01-05 10:17:11 -07:00
Ben Johnson	dcd2a771b0	Optimizing tsi compaction.	2017-01-05 10:17:11 -07:00
Ben Johnson	1ce99e797f	Use series map in tsi1.LogFile.	2017-01-05 10:17:11 -07:00
Ben Johnson	9b1e8215e0	Remove dictionary encoding, add bulk series insertion.	2017-01-05 10:17:11 -07:00
Ben Johnson	5f7654173e	Add locking to sketch merge.	2017-01-05 10:17:11 -07:00
Ben Johnson	9bd19cdc69	Fix inmem DELETE SERIES.	2017-01-05 10:17:11 -07:00
Ben Johnson	f9efcb3365	Re-add shared in-memory index.	2017-01-05 10:17:09 -07:00
Edd Robinson	0f9b2bfe6a	Fix tests	2017-01-05 10:16:15 -07:00
Edd Robinson	4ccb8dbab1	Move series count check to shard	2017-01-05 10:16:13 -07:00
Edd Robinson	0cb74eedbf	Add log file (WAL) sketches	2017-01-05 10:15:38 -07:00
Edd Robinson	190c78c644	Add series sketches	2017-01-05 10:15:37 -07:00
Edd Robinson	695adafc00	Add measurement sketches	2017-01-05 10:15:37 -07:00
Ben Johnson	745b1973a8	tsi compaction	2017-01-05 10:15:37 -07:00
Ben Johnson	83e80f6d0b	Fix in-mem index integration tests.	2017-01-05 10:15:37 -07:00
Ben Johnson	183418dcbd	Fix tsi TAG KEYS iterator.	2017-01-05 10:15:36 -07:00
Ben Johnson	759ff4ab80	Add tsi1 term hash index.	2017-01-05 10:15:35 -07:00
Ben Johnson	75cfe244c4	Add series hash index.	2017-01-05 10:15:35 -07:00
Ben Johnson	9f8b206b51	Fix measurement system queries.	2017-01-05 10:15:34 -07:00
Ben Johnson	4aa78383d1	Fix tsi1 series deletion.	2017-01-05 10:14:48 -07:00
Ben Johnson	5965610de6	Refactoring tsi tombstoning.	2017-01-05 10:14:02 -07:00
Ben Johnson	e7940cc556	Add tsi1 series system iterator.	2017-01-05 10:14:00 -07:00
Ben Johnson	87f4e0ec0a	Add regex support in tsi1.	2017-01-05 10:12:29 -07:00
Ben Johnson	d13afa8f47	Iterator refactoring	2017-01-05 10:11:49 -07:00
Jason Wilder	1ba64f3610	Disable max-value-per-tag option temporarily This is too slow currently and causes all writes to timeout.	2017-01-05 10:11:47 -07:00
Jason Wilder	f0427d180e	Fix tsi index panics Hardcoded panics cause the server to crash in 10s due to stats collection.	2017-01-05 10:11:12 -07:00
Jason Wilder	4bf7b2bb19	Allow tsi to be enabled via config option	2017-01-05 10:11:12 -07:00
Jason Wilder	2b96c5d4d0	Set Tags on entry These were lost when reloading from the index. Fixes queries not returning any data.	2017-01-05 10:11:12 -07:00
Jason Wilder	a6490920fd	Fix reslicing indices The slicing was backwards causing the buffer to grow indefinitely and filling the disks on writes.	2017-01-05 10:11:12 -07:00
Jason Wilder	59864226b7	Add RWMutex to LogFile Fixes concurrent map access panic	2017-01-05 10:11:12 -07:00
Ben Johnson	fbe7f464ee	Improve insert performance.	2017-01-05 10:11:12 -07:00
Ben Johnson	bf89b94d17	Fix WalkTagKeys().	2017-01-05 10:11:11 -07:00
Ben Johnson	33412782ed	Fix go vet issue.	2017-01-05 10:11:10 -07:00
Ben Johnson	2b864c72c5	Refactor MeasurementBlockTrailer read/write.	2017-01-05 10:11:10 -07:00
Ben Johnson	cb93f10120	Remove per-shard in-memory index.	2017-01-05 10:11:09 -07:00
Ben Johnson	409b0165f5	shared in-memory index	2017-01-05 10:09:57 -07:00
Ben Johnson	a812502ea3	reintegrating in-memory index	2017-01-05 10:07:35 -07:00
Ben Johnson	1ac067e53b	intermediate	2017-01-05 10:03:09 -07:00
Ben Johnson	fda84955ea	Remove TODO	2017-01-05 10:02:42 -07:00
Ben Johnson	62d2b3ebe9	Series filtering.	2017-01-05 10:02:42 -07:00
Ben Johnson	62269c3cea	intermediate	2017-01-05 10:02:41 -07:00
Ben Johnson	5f5b02e052	intermediate	2017-01-05 10:01:49 -07:00
Ben Johnson	8863e3c0f3	Refactor tsi1 merge iterators, finish multi-file compaction.	2017-01-05 10:01:25 -07:00
Ben Johnson	e3af4b0dad	Refactor iterators.	2017-01-05 10:00:45 -07:00
Ben Johnson	ce9e3181a5	Refactor merge iterators.	2017-01-05 10:00:45 -07:00
Ben Johnson	0294e717a0	Add mm, tag key, tag value, & series iterators.	2017-01-05 10:00:44 -07:00
Ben Johnson	2bfafaed76	tsi1 log compaction	2017-01-05 10:00:44 -07:00
Ben Johnson	afce53e81b	Rebase fixes.	2017-01-05 10:00:44 -07:00
Ben Johnson	992e651588	Add tsi1.Log.	2017-01-05 10:00:44 -07:00
Ben Johnson	2a81351992	Implement tsdb.Index interface on tsi1.Index.	2017-01-05 10:00:43 -07:00
Edd Robinson	e2c3b52ca4	Adds a custom HyperLogLog++ implementation	2017-01-05 10:00:14 -07:00
Edd Robinson	da63b349a4	Fix bad rebase	2017-01-05 09:59:44 -07:00
Edd Robinson	ebc92ca04f	Fix overflow issues	2017-01-05 09:59:12 -07:00
Edd Robinson	149b1cef1d	Fix 32bit overflow; limit capacity	2017-01-05 09:59:10 -07:00
Edd Robinson	33623c1fa9	Revert back to original approach	2017-01-05 09:58:39 -07:00
Edd Robinson	9ed6040265	Tidy up	2017-01-05 09:58:37 -07:00
Edd Robinson	2a5c865b44	Use xxhash	2017-01-05 09:57:35 -07:00
Edd Robinson	2d9bd09784	Use []byte where possible in Index	2017-01-05 09:57:34 -07:00
Edd Robinson	3edbfb9197	Prevent panic when shard nil	2017-01-05 09:56:51 -07:00
Edd Robinson	3187cd4432	Cleanup series created stat	2017-01-05 09:56:49 -07:00
Edd Robinson	4b1ef68dc9	Move series and measurement stats to store	2017-01-05 09:54:05 -07:00
Edd Robinson	aaf85ae38d	Tombstoning with series cardinality part 1	2017-01-05 09:54:04 -07:00
Edd Robinson	bd8dd9a291	Sketches working	2017-01-05 09:54:04 -07:00
Edd Robinson	d19fbf5ab4	Wire in HLL estimator	2017-01-05 09:54:03 -07:00
Edd Robinson	2b8efefef4	Initial index interface	2017-01-05 09:51:43 -07:00
Edd Robinson	05bc4dec00	Refactor	2017-01-05 09:50:23 -07:00
Edd Robinson	c535e3899a	Remove in-memory index from Shard and Store	2017-01-05 09:47:09 -07:00
Edd Robinson	2171d9471b	Initialise index in shards	2017-01-05 09:42:48 -07:00
Ben Johnson	57d0556174	Fix 32-bit issues.	2017-01-05 09:34:37 -07:00
Ben Johnson	41f2babe66	Minor TSI index benchmark refactor	2017-01-05 09:34:37 -07:00
Ben Johnson	ac9c6a0207	Add TSI index benchmark.	2017-01-05 09:34:37 -07:00
Ben Johnson	8d40ceb00c	TSI1 Index	2017-01-05 09:34:36 -07:00
Ben Johnson	9b62df23d2	Add MeasurementBlock.	2017-01-05 09:34:36 -07:00
Ben Johnson	3240af07e0	Fix RHH packing.	2017-01-05 09:34:36 -07:00
Ben Johnson	e25d61e4bd	TagSet writer & reader.	2017-01-05 09:34:36 -07:00
Ben Johnson	4eeb81ef38	Add SeriesList tombstoning.	2017-01-05 09:34:36 -07:00
Ben Johnson	2c34b24f5c	Implemented SeriesList	2017-01-05 09:34:36 -07:00
Ben Johnson	6523675c20	Implemented RHH hash map.	2017-01-05 09:34:35 -07:00
Mark Rushakoff	6a94d200c8	Merge remote-tracking branch 'influx/master' into mr-godoc	2017-01-04 13:27:36 -08:00
Mark Rushakoff	89a587e865	Use one atomic operation in (Cache).decreaseSize The previous implementation was susceptible to a race condition (of correctness) since c.decreaseSize is called without a lock in (Cache).WriteMulti. There were already tests which asserted the correctness of the result of decreaseSize, so no tests were added or modified.	2017-01-04 13:13:31 -08:00
Cory LaNou	3c518f8927	panicing is bad -> error returns are good	2017-01-03 14:28:29 -06:00
Mark Rushakoff	07b87f2630	Miscellaneous lint cleanup	2017-01-03 09:47:32 -08:00
Mark Rushakoff	41415cf2fb	Update godoc for tsm1 package	2017-01-02 07:30:18 -08:00
Mark Rushakoff	4a774eb600	Update godoc for the tsdb package	2016-12-30 21:12:37 -08:00
Gustav Westling	26b33307ae	Resolved PR comments on test files	2016-12-30 11:42:38 +01:00
Gustav Westling	56d98325da	Removed ineffective assignments, and added checks for errors that previsouly was not checked	2016-12-29 20:26:15 +01:00
Jason Wilder	2468347ffb	Fix comment	2016-12-19 14:17:49 -07:00
Jason Wilder	326557e539	Fix race in partition.reset	2016-12-19 14:17:01 -07:00
Jason Wilder	e91e45d71c	Fix panic in cache benchmark	2016-12-19 14:17:01 -07:00
Jason Wilder	0b6b9ea1cb	Use atomics for cache.snapshotSize stat	2016-12-19 14:17:01 -07:00
Jason Wilder	637a67ea35	Reduce lock contention on measurementFields	2016-12-19 14:17:01 -07:00
Jason Wilder	b7c1e625b0	Move needSort tracking to Deduplicate This eliminates some *UnixNano() calls and also simplifies the cache logic so that it does not need to worry about whether entries are sorted.	2016-12-19 14:17:01 -07:00
Jason Wilder	dea87703cd	Reduce UnixNano pointer call	2016-12-19 14:17:01 -07:00
Mark Rushakoff	722b6345fe	Fix unchecked error in templated Read${TYPE}Block	2016-12-19 09:31:26 -08:00
Jonathan A. Sternberg	ec57108520	Use proper uber-go/zap import path It looks like the real import path to the project is go.uber.org/zap instead of github.com/uber-go/zap since the example in the project references that path.	2016-12-15 08:54:14 -06:00
Edd Robinson	ec27c57127	Further optimisations and a race fix	2016-12-14 18:23:36 +00:00
Edd Robinson	05ec6ad9ad	Add to index safely	2016-12-14 18:23:36 +00:00
Edd Robinson	d78ca1a0f3	Fix some races	2016-12-14 18:23:36 +00:00
Edd Robinson	d2923c7bf9	Add hints as to how to pre-allocate entry values Currently, whenever a snapshot occurs the Cache is reset and so many allocations are repeated, as the same type of data is re-added to the Cache. This commit allows the stores to keep track of the number of values within an entry, and use that size as a hint when the same entry needs to be recreated after a snapshot. To avoid hints persisting over a long period of time they are deleting after every snapshot, and rebuilt using the most recent entries only.	2016-12-14 18:23:36 +00:00
Edd Robinson	f2b5c7f5be	Reduce contention when adding entries	2016-12-14 18:23:36 +00:00
Edd Robinson	98f0392ca6	Update size using atomic	2016-12-14 18:23:36 +00:00
Edd Robinson	66edb32182	Sharded Cache using a hash ring	2016-12-14 18:23:36 +00:00
Edd Robinson	d3e6d4e7ca	Add benchmarks	2016-12-14 18:21:50 +00:00
Jonathan A. Sternberg	21502a39e8	Switch logging to use structured logging everywhere The logging library has been switched to use uber-go/zap. While the logging has been changed to use structured logging, this commit does not change any of the logging statements to take advantage of the new structured log or new log levels. Those changes will come in future commits.	2016-12-14 10:45:15 -06:00
gunnaraasen	78b1a0e771	Add stats on dropped measurements and series; Fixes #7697	2016-12-13 15:17:31 -08:00
Jason Wilder	4f28c90b54	Optimize Value.Deduplicate Deduplicate is called from various places in the engine and can cause a lot of garbage to get created. It first creates a map and then adds each value to the map in order (1st alloc). It then creates a new slice (2nd alloc) and appends everything from the map to the slice. Finally, it sorted the new slice (3rd alloc). This switches the algorithm to use stable sorting and resuing the existing slice to avoid allocations.	2016-12-08 21:10:56 -07:00
Hrvoje Marjanovic	9483b8b409	gofmt	2016-12-03 22:06:38 +01:00
Hrvoje Marjanovic	6ed708e3fd	Reduce pool size, change WAL writers default Big pool can lead to huge memory usage in certain loads. See #7640 for detailed discussion.	2016-12-02 18:45:43 +01:00
Allen Petersen	31129ab0e9	Use slash separator for filenames in tar archives NO-OP on platforms with unix path separator. On Windows paths get converted to slashes before adding to archive and back to backslashes during restore.	2016-11-29 09:44:08 -08:00
Jason Wilder	27d157763a	Merge pull request #7651 from influxdata/jw-shard-last-modified Expose Shard.LastModified	2016-11-23 10:19:26 -07:00
Jason Wilder	e8a28cfbab	Expose Shard.LastModified This returns the LastModified time of the shard. The LastModified time is the wall time when a change to the shards state occurred. It uses the WAL or FileStore to determine the max mod time.	2016-11-23 10:04:07 -07:00
Edd Robinson	b83b8df32f	Merge pull request #7635 from influxdata/er-msg Fix incorrect error message	2016-11-23 13:58:33 +00:00
Edd Robinson	9e9719749f	Sprinkle some golint	2016-11-17 16:31:38 +00:00
Edd Robinson	28ba8ced74	Fixes #7625	2016-11-17 16:31:36 +00:00
Jason Wilder	3a5a01181b	Switch all Value types from pointers	2016-11-15 16:13:55 -07:00
Jason Wilder	bf17074f58	Avoid allocation when counting tag keys A new sorted slice was called by the monitor func every 10s. The tag keys don't need to be sorted so this avoid the allocation of the slice and one during sorting.	2016-11-15 16:13:55 -07:00
Jason Wilder	0ee58c208a	Switch time.Sleep to time.Ticker Avoids an allocation when calling time.Sleep	2016-11-15 16:13:55 -07:00
Jason Wilder	73b8f52ca0	Cache results onf findGenerations This allocates quite a bit and it's called multiple times per second per shard. The generations don't change until a compaction has occurred so most of the time is re-calculating the same thing and creating garbage.	2016-11-15 16:13:55 -07:00
Jason Wilder	0b6f5441b9	Add config option to messages when limits exceeded When a limit is exceeded, we return errors and sometimes log (if appropriate) that a limit was exceeded. The messages don't always provide an indication as to where or how they are configured. Instead, return the config option (easily searchable for) as well as the limit currently set and the value that exceeded it when possible.	2016-10-28 14:54:45 -06:00
Jason Wilder	b1ceb5e66d	Add cache write OK, Dropped, Error stats Adds a new dropped stat as well as fixes OK and error stats not actually get collected and stored.	2016-10-28 12:15:50 -06:00
Jason Wilder	873189e0c2	Fix panic: interface conversion: tsm1.Value is tsm1.FloatValue, not tsm1.StringValue If concurrent writes to the same shard occur, it's possible for different types to be added to the cache for the same series. The way the measurementFields map on the shard is updated is racy in this scenario which would normally prevent this from occurring. When this occurs, the snapshot compaction panics because it can't encode different types in the same series. To prevent this, we have the cache return an error a different type is added to existing values in the cache. Fixes #7498	2016-10-28 12:15:50 -06:00
Jason Wilder	e388912b6c	Fix race in findGenerations The file store stats slice is re-used which causes the race below: WARNING: DATA RACE Write at 0x00c42007e140 by goroutine 43: github.com/influxdata/influxdb/tsdb/engine/tsm1.(FileStore).Stats() /Users/jason/go/src/github.com/influxdata/influxdb/tsdb/engine/tsm1/file_store.go:511 +0x22e github.com/influxdata/influxdb/tsdb/engine/tsm1.(DefaultPlanner).findGenerations() /Users/jason/go/src/github.com/influxdata/influxdb/tsdb/engine/tsm1/compact.go:461 +0x6f github.com/influxdata/influxdb/tsdb/engine/tsm1.(DefaultPlanner).PlanLevel() Previous read at 0x00c42007e140 by goroutine 40: github.com/influxdata/influxdb/tsdb/engine/tsm1.(DefaultPlanner).findGenerations() /Users/jason/go/src/github.com/influxdata/influxdb/tsdb/engine/tsm1/compact.go:463 +0x13d github.com/influxdata/influxdb/tsdb/engine/tsm1.(*DefaultPlanner).PlanOptimize()	2016-10-28 12:15:49 -06:00
Jason Wilder	96c9fb3648	Actually update the defaults for TSM 7510 update the defaults in the sample config, but did not update the code. This updates the defaults in the config that changed.	2016-10-26 09:49:25 -06:00
Steven Hartland	3f16197243	Improve tsm1 cache performance Reduce the cache lock contention by widening the cache lock scope in WriteMulti, while this sounds counter intuitive it was: * 1 x Read Lock to read the size * 1 x Read Lock per values * 1 x Write Lock per values on race * 1 x Write Lock to update the size We now have: * 1 x Write Lock This also reduces contention on the entries Values lock too as we have the global cache lock. Move the calculation of the added size before taking the lock as it takes time and doesn't need the lock. This also fixes a race in WriteMulti due to the lock not being held across the entire operation, which could cause the cache size to have an invalid value if Snapshot has been run in the between the addition of the values and the size update. Fix the cache benchmark which where benchmarking the creation of the cache not its operation and add a parallel test for more real world scenario, however this could still be improved. Add a fast path newEntryValues values for the new case which avoids taking the values lock and all the other calculations. Drop the lock before performing the sort in Cache.Keys().	2016-10-25 15:24:51 -06:00
Jonathan A. Sternberg	a515aeda39	Optimize first/last when no group by interval is present The `first()` and `last()` functions response rate would increase linear to the number of points even though it seems like it shouldn't. This optimization greatly reduces the amount of time to return a response when no `GROUP BY time(...)` clause is present in a query.	2016-10-25 09:57:31 -05:00
Jason Wilder	686d1a7ba4	Remove unused config options	2016-10-24 15:32:38 -06:00
Edd Robinson	0ee093f1fb	Memoize output of FileStore.Stats	2016-10-24 10:23:20 -06:00
Jonathan A. Sternberg	3681bc8a43	Filter out series within shards that do not have data for that series Previously, we would return a full tag set for every shard and the tag set would include all series that existed in the database index including series that didn't physically exist within that shard. This led to the tag sets returned being incredibly huge when we had high cardinality but sparse data. Since the data was sparse, it was unexpected that it would cause such a large strain on the system by most people. Now we filter out the series ids that are not assigned to the current shard when computing a tag set for that shard. This lowers the memory usage for high cardinality sparse data drastically and allows queries on those to complete successfully. This does not resolve issues for high cardinality data in every shard that is also spread out over a long series of time. That situation isn't nearly as common as the above situation though.	2016-10-20 14:15:34 -05:00
Jason Wilder	2e473e9518	Fix panic in AppendSeriesKeyByID Calling this function with a series ID that does not exist in the measurement causes a panic. Fixes #7334	2016-10-19 11:07:19 -06:00
Jason Wilder	b50d9558cf	Merge pull request #7479 from influxdata/jw-clean-err Skip cleanup if dir does not exist	2016-10-18 15:49:09 -06:00
Jason Wilder	f30b00c24f	Skip cleanup if dir does not exist	2016-10-18 15:33:39 -06:00
Mark Rushakoff	377c40f122	Add stats for active compactions Unify logic around compaction execution to a single place. Also report on the error stats that we track. Previously they were not emitted in the stats output.	2016-10-18 14:12:21 -07:00
Joe LeGasse	de9c743004	TSM: update comments for disabling level compactions	2016-10-18 14:14:59 -06:00
Joe LeGasse	eda8f70372	TSM: Handle concurrent deletes for compaction	2016-10-18 14:14:59 -06:00
Jason Wilder	47b8049e48	Update comment	2016-10-18 14:14:53 -06:00
Jason Wilder	ed7975874f	Rename Enabled -> Enable	2016-10-18 12:22:00 -06:00
Jason Wilder	f254b4f3ae	Allow snapshot compactions during deletes If a delete takes a long time to process while writes to the shard are occuring, it was possible for the cache to fill up and writes to be rejected. This occurred because we disabled all compactions while writing tombstone file to prevent deleted data from re-appearing after a compaction completed. Instead, we only disable the level compactions and allow snapshot compactions to continue. Snapshots already handle deleted data with the cache and wal. Fixes #7161	2016-10-18 12:14:51 -06:00
Jonathan A. Sternberg	41e4e73d4e	Reduce map allocations when computing the TagSets of a measurement Instead of assigning a boolean value of true to the filter expressions when there was no meaningful expression, this drops a boolean expression of true from the filter expressions so we don't have to perform a map assignment. This allows us to reduce allocations and assignments when a `WHERE` clause only contains tag comparisons and no field comparisons.	2016-10-17 12:13:19 -05:00
Jason Wilder	a5f871d62c	Rework monitoring to avoid allocations	2016-10-10 11:42:15 -06:00
Jason Wilder	bbecb3f03d	Drop points that would execeed limits This changes the behavior of the max-series-per-database and max-values-per-tag limits to drop points that would exceed the limits and allow the remaining points to be written. Previously, the whole batch would fail and return and 500 error to the client. This now will write the allow points and return a `partial write` error indicating some of the points were dropped, how many were dropped and one of the problem measureent and tags.	2016-10-10 11:42:15 -06:00
Jason Wilder	8fce6bba48	Add tag value cardinality limit	2016-10-10 11:42:15 -06:00
Mark Rushakoff	5ae8cf8312	Speed up shutdown On my machine with about 20 shards, it would take 10+ seconds to shut down InfluxDB with SIGINT. After this change, it shuts down in nearly instantly. (tsdb.Store).Close was shutting down each of its shards sequentially. Each shard's engine would signal to its compaction goroutines to quit, and because each compaction goroutine has a hardcoded 1-second sleep in between checks, waiting for the goroutines would often block for up to a second. This change closes all of the TSDB store's shards in parallel. This means it's possible that multiple close values could error at once, but we're still only returning the first error, consistent with previous behavior. That being said, the return value of (tsdb.Store).Close is ignored in (*cmd/influxd/run.Server).Close anyway.	2016-10-10 09:18:47 -07:00
Jason Wilder	798fa0a9f8	Return error with unknown field type This will just panic when trying to snapshot the value because EmptyValue can't be written to TSM files.	2016-10-03 16:30:21 -06:00
Jason Wilder	125f106956	Pre-size the values map when write points	2016-10-03 16:30:21 -06:00
Joe LeGasse	743946fafb	models: Add FieldIterator type The FieldIterator is used to scan over the fields of a point, providing information, and delaying parsing/decoding the value until it is needed. This change uses this new type to avoid the allocation of a map for the fields which is then thrown away as soon as the points get converted into columns within the datastore.	2016-10-03 16:30:21 -06:00
Jason Wilder	20f1fb3f7f	Replace gotos with anonymous functions	2016-10-03 12:08:53 -06:00
Jason Wilder	750c8b3932	Reduce lock contention in cache.Values The cache read lock was held for the whole duration of the call when it only needs to be held at the beginning since entries have their own locks.	2016-10-03 10:21:54 -06:00
Jason Wilder	1b462312a9	Re-use decoder pools The decoders were held onto each iterator to avoid creating them all the time. Some of them have use quite a bit of memory so they can be expensive to create when querying across many series. Intead, more them to a re-usable pool where we create the minimum that could active be in use. This reduces garbage as well as makes the iterators less expensive to create.	2016-10-03 10:21:54 -06:00
Jason Wilder	f727effd7f	Merge pull request #7385 from influxdata/jw-query-allocs Reduce query planning allocations	2016-10-03 09:08:36 -06:00
Jason Wilder	a15a416eaa	Fix decoding RLE integer blocks with negative deltas Integer blocks that were run length encoded could produce the wrong value when read back out because the deltas were not zig zag decoded before scaling the final value. If the deltas were negative, as would be seen in a counter that decrements by a constant value, the results would be random with som negative and positive values. Fixes #7391	2016-10-02 23:51:29 -06:00
Jason Wilder	68dd312bb1	Reduce allocations when calculating tagsets The TagSets function was creating a lot of intermediate maps and slices to calculate the sorted tag sets. It first creates a map to group tag sets with their series, it then created an equally sized slice of the tag keys and sorted then. Finally, it created a new slice and added the tag sets in the original map by the ordering of the sorted keys. It was also recreating the tags map multiple time creating extra garbage in the loop. This simplifies the code to create one map for grouping and than adding the distinct sets to a slice which is then sorted. It also fixes the multple tag maps getting created.	2016-09-29 16:02:29 -06:00
Mark Rushakoff	97c2f6f5c1	Add walPath tag to shard stats Without the WAL path as a tag, the diskBytes field looked like it was reporting the size of the data directory incorrectly. Fixes #7382.	2016-09-29 10:19:11 -07:00
Jason Wilder	dcb65865a2	Merge pull request #7376 from influxdata/jw-revert Revert re-using byte slices during compactions	2016-09-28 08:24:35 -06:00
joelegasse	87ecd97e7b	Merge pull request #7371 from influxdata/2016-09-27--rw--use-gotos-for-encoding-cleanup Gotos to simplify uses of the new encoder pools.	2016-09-28 08:57:33 -04:00
Jason Wilder	1755f20d2a	Revent re-using byte slices during compactions This is causing a fatal error: fault panic when packing blocks.	2016-09-27 23:41:06 -06:00
Jonathan A. Sternberg	e22e33d5fd	Merge pull request #7374 from influxdata/merge-from-1.0.1 Merge tag 'v1.0.1'	2016-09-27 20:32:58 -05:00
Jonathan A. Sternberg	3afdf3cd94	Merge tag 'v1.0.1'	2016-09-27 17:53:33 -05:00
rw	c3fc87b619	Remove dangling named return value.	2016-09-27 14:18:32 -07:00
rw	fcd425c8c6	Incorporate style feedback from Joe.	2016-09-27 14:07:06 -07:00
rw	47c1c6763c	Use encoder reset to save on allocs.	2016-09-27 13:31:35 -07:00
rw	9429a2f96a	Gotos to simplify uses of the new encoder pools. For maintainability.	2016-09-27 11:47:25 -07:00
Jason Wilder	5367372253	Merge pull request #7364 from influxdata/2016-09-26-fix-data-race-in-write-path Fix data race in *tsdb.Shard write path.	2016-09-26 18:34:19 -06:00
rw	f131d3cc77	Fix off-by-one error that could panic.	2016-09-26 17:03:03 -07:00
rw	3e0d3be461	Use pre-existing function.	2016-09-26 13:12:10 -07:00
rw	bea010b5f3	Fix data race in *tsdb.Shard write path. Ensure that the Shard's Index is read-locked before calculating the count of its constituent series.	2016-09-26 12:42:35 -07:00

... 2 3 4 5 6 ...

1631 Commits (40ec85aacd07069e4815c33d051d646411d0076e)