influxdb

Commit Graph

Author	SHA1	Message	Date
Ben Johnson	45f1c28adb	add tsm iterator stats buffer This commit adds a buffer for stats to be updated without requiring a mutex lock/unlock on every point. The tradeoff is that stats are not exactly precise. This works for our use case because stats are only periodically checked.	2016-03-23 12:23:22 -06:00
Jonathan A. Sternberg	a35d9602cd	Fix where filters when a OR is used and when a tag does not exist If an OR was used, merging filters between different expressions would not work correctly. If one of the sides had a set of series ids with a condition and the other side had no series ids associated with the expression, all of the series from the side with a condition would have the condition ignored. Instead of defaulting a non-existant series filter to true, it should just be false and the evaluation of the one side that does exist should take care of determining if the series id should be included or not. The AND condition used false correctly so did not have to be changed. If a tag did not exist and `!=` or `!~` were used, it would return false even though the neither a field or a tag equaled those values. This has now been modified to correctly return the correct series ids and the correct condition. Also fixed a panic that would occur when a tag caused a field access to become unnecessary. The filter using the field access still got created and used even though it was unnecessary, resulting in an attempted access to a non-initialized map. Fixes #5152 and a bunch of other miscellaneous issues.	2016-03-22 12:19:06 -04:00
Ben Johnson	573dd0f96a	Merge pull request #6035 from benbjohnson/query-engine-reduce-alloc Reduce allocations in query execution	2016-03-22 10:11:14 -06:00
Ben Johnson	6e1c1da25b	reduce allocations in query execution This commit removes some heap objects by converting them from pointer references to non-pointers or by reusing buffers.	2016-03-22 09:51:39 -06:00
Jason Wilder	7857e07a1e	Merge pull request #6062 from influxdata/mr-prune-wal-config Remove unused WAL configuration variables/fields	2016-03-22 09:20:27 -06:00
Jonathan A. Sternberg	ad96207868	Fix ORDER BY desc so it doesn't skip values After reading the initial buffer, ORDER BY desc would read the next block into the buffer and only read the first element. It's because the code that was copied from the ascending cursor wasn't modified correctly to set the position to the last element in the buffer. The buffer size has also been lowered from 1000 to 10 to match with the ascending cursor for performance with limit queries. Fixes #6055.	2016-03-22 09:40:11 -04:00
Ben Johnson	7156c1f9bd	add IteratorStats This commit adds an `IteratorStats` that holds aggregate iterator processing information. A method is also added to `Iterator` to return the stats: Stats() influxql.IteratorStats The remote iterators will also emit their stats in the point stream upon first connection, on a given interval, and then finally once the last point has been sent.	2016-03-21 16:25:19 -06:00
Jason Wilder	ee2f21e76f	Merge pull request #6082 from influxdata/jw-tsm Fix partially written TSM files	2016-03-21 15:42:27 -06:00
Jason Wilder	7567453c9a	Ensure TSM files are fsync'd Make sure TSM files are fsync'd when closed and also that the parent dir is fsync'd when they are renamed.	2016-03-21 15:03:52 -06:00
Jason Wilder	a4e5446ddd	Return error when TSM writer close returns one The TSM writer uses a bufio.Writer that needs to be flushed before it's closed. If the flush fails for some reason, the error is not handled by the defer and the compactor continues on as if all is good. This can create files with truncated indexes or zero-length TSM files. Fixes #5889	2016-03-21 15:00:36 -06:00
Jonathan A. Sternberg	6655ca7769	Create a new interrupt iterator that will stop emitting points after an interrupt Use of the iterator is spread out into both `IteratorCreators` and inside of the iterators themselves. Part of the interrupt must be handled inside of the engine so it stops trying to emit points when an interrupt is found and another part of the interrupt has to happen when combining the iterators so it doesn't just start reading the next shard.	2016-03-21 12:07:07 -04:00
Mark Rushakoff	7a2adfcc5d	Remove unused WAL configuration variables/fields These were all b1/bz1 settings that no longer have any effect: - {Default,}MaxWALSize - {Default,}WALFlushInterval - {Default,}WALPartitionFlushDelay - {Default,WAL}ReadySeriesSize - {Default,WAL}CompactionThreshold - {Default,WAL}MaxSeriesSize - {Default,WAL}FlushColdInterval - {Default,WAL}PartitionSizeThreshold	2016-03-20 13:16:52 -07:00
Jonathan A. Sternberg	d75428f79f	Rename the special condition "name" to "_name" to reduce conflicts Fixes #6034.	2016-03-16 17:17:04 -04:00
Jonathan A. Sternberg	eb2d49dbe4	Merge pull request #6007 from benbjohnson/explicit-system-names Allow querying of system-like series	2016-03-15 16:15:17 -04:00
Cory LaNou	ba6a95e9bc	Merge pull request #5994 from influxdata/single-server-lite Single Server	2016-03-14 16:11:37 -05:00
Ben Johnson	f692621ef5	allow querying of system-like series Internal system series start with an underscore prefix but restricting this prevents users who already use an underscore prefix in their series names. Fixes #5870	2016-03-14 13:50:52 -06:00
Jason Wilder	3fd40d48a1	Merge pull request #6006 from influxdata/jw-deadlock Fix deadlock when running backup	2016-03-14 13:36:45 -06:00
Jason Wilder	9984cd5d6d	Fix skipping blocks at query time when overlaps exist Depending on how data is written across TSM files, it was possible to skip over some blocks at query time making it looks like data was missing.	2016-03-14 13:11:11 -06:00
Jason Wilder	000459e350	Fix deadlock when running backup A deadlock occurs under write load if a backup is run in between the time when a snapshot compactions has snapshotted the cache and successfully written it to disk. The issus is that the second snapshot call will block on the commit lock while it is holding the engine write lock. This causes all writes to block as well as prevents the currently runnign snapshot compaction from completing because it needs to acquire a read-lock. This PR removes the commit lock and just returns an error if a snapshot is in progress to all any locks being held to be released. The caller can determine whether to retry or giveup.	2016-03-14 12:36:48 -06:00
Cory LaNou	27cfaa4b7a	in memory meta, single node configs, etc.	2016-03-14 16:55:54 +00:00
Joe LeGasse	344e5abd41	Changed type-switch a few places to reduce allocations. Slices of tsm1.Value interfaces are only ever used with all the same types, and the previous code would switch on the type returned from a call to Value(), which allocated and returned an interface{} object for the underlying value. This change instead type-switches on the tsm1.Value object itself, allowing it direct access to the underlying value field, eliminating the unecessary allocations.	2016-03-11 15:57:05 -05:00
Ben Johnson	beda072426	add support for remote expansion of regex This commit moves the `tsdb.Store.ExpandSources()` function onto the `influxql.IteratorCreator` and provides support for issuing source expansion across a cluster.	2016-03-11 12:40:07 -07:00
Jason Wilder	c44195d999	Convert measurementToRegex to exported func Make it consistent with other conventions where exported funcs take a lock.	2016-03-09 17:45:37 -07:00
Jason Wilder	992c78ee22	Remove period shard maintenance goroutine This is no longer used in tsm and just peridocially locks everything for no reason now.	2016-03-09 17:31:02 -07:00
Jason Wilder	ae2360df7c	Use read lock to expand sources A write-lock was taken which locks the whole store during a query that needs to expand sources. Under load, writes can start to fail.	2016-03-09 17:22:57 -07:00
Edd Robinson	7dbc0f49d3	Merge pull request #5818 from influxdata/er-upgrade-error Highlight upgrade info for old shards	2016-03-09 19:39:59 +00:00
Edd Robinson	58c03448aa	Merge pull request #5514 from influxdata/er-engine-panic Ensure shards and engine are safely closed	2016-03-09 18:56:36 +00:00
Ben Johnson	41dde61226	SHOW SERIES	2016-03-08 11:47:57 -07:00
Jonathan A. Sternberg	2f0e246757	Implemented the tag values iterator for `SHOW TAG VALUES` `SHOW TAG VALUES` output has been modified to print the measurement name for every measurement and to return the output in two columns: key and value. An example output might be: > SHOW TAG VALUES WITH KEY IN (host, region) name: cpu --------- key value host server01 region useast name: mem --------- key value host server02 region useast `measurementsByExpr` has been taught how to handle reserved keys (ones with an underscore at the beginning) to allow reusing that function and skipping over expressions that don't matter to the call. Fixes #5593.	2016-03-06 09:52:34 -05:00
Jason Wilder	e3fef5593c	Merge pull request #5855 from jonseymour/jss-5854-go-master-breaks-build fix tests to cope with future changes to testing.quick.Check - see #5854	2016-03-01 19:03:21 -07:00
Mark Rushakoff	cdcb079769	Tag TSM stats with database, retention policy ... by extracting the db/rp from the given path. Now that the code has "standardized" on extracting db/rp this way, the ShardLocation struct is no longer necessary and thus has been removed. We're back on the previous style of passing the path and walPath to NewShard.	2016-02-29 09:17:34 -08:00
Jon Seymour	73b3a2a056	Merge #5855 (issue: #5854 ). RHS merges cleanly with 0.10.0 Signed-off-by: Jon Seymour <jon@wildducktheories.com>	2016-02-29 20:37:32 +11:00
Jon Seymour	716cdd7f41	tsm: modify encoding tests to deal with possible nil slices from testing.quick.Check in go master The current go compiler at the tip of the go master (1d5001af) has a modified implementation of testing.quick.Check that now generates nil slices as test data. (See: https://gophers.slack.com/archives/general/p14567053570110). The existing tests expect round tripping to work in this case but it does not. So, in these cases we change the expectation to reflect actual behaviour. This needs to be checked for reasonableness.	2016-02-29 20:36:19 +11:00
Jonathan A. Sternberg	aa0b603938	Convert `SHOW FIELD KEYS` to the new query engine Fixes #5579.	2016-02-25 18:31:02 -05:00
Jason Wilder	8d70d65a82	Convert time.Time to int64	2016-02-25 15:15:01 -07:00
Jason Wilder	55a503671d	Merge pull request #5833 from jonseymour/jss-5832-snapshot-may-not-be-sorted tsm: cache: need to check that snapshot has been sorted	2016-02-25 15:10:49 -07:00
Mark Rushakoff	40a98e0d55	Add database, RP as tags on shard stats This commit updates tsdb.Shard to contain a ShardConfig and updates tsdb.Store to directly reference a map of tsdb.Shard rather than the previous tsdb.shardLocation abstraction.	2016-02-25 13:41:55 -08:00
Jon Seymour	11123d2694	Merge #5833 (issue: #5832 ). Signed-off-by: Jon Seymour <jon@wildducktheories.com>	2016-02-26 07:59:03 +11:00
Jon Seymour	2c7cd06b99	tsm: cache: need to check that snapshot has been sorted. Previously, the for loop at the end of the method assumed that all entries had been deduplicated, including the entry discovered in the snapshot. However, this wasn't actually true. With this change, we make it true. Signed-off-by: Jon Seymour <jon@wildducktheories.com>	2016-02-26 07:56:25 +11:00
Jon Seymour	7eabae68de	tsm: cache: add a test for the write sequence {6,1,snapshot,7,2} Consider the write sequence: 6,1,snapshot,7,2. The hot cache gets deduplicated, so is 2,7. Now consider the test if 1 >= 2, this is false, so needSort is not set to true. The problem is the implicit assumption that the snapshot is always sorted by the time that merged() runs, but this may not be true. Signed-off-by: Jon Seymour <jon@wildducktheories.com>	2016-02-26 07:43:50 +11:00
Jason Wilder	6ebc192298	Merge pull request #5678 from jonseymour/typo doc: typographical, spelling, grammar, word-choice and phrasing improvements.	2016-02-25 09:33:41 -07:00
Jason Wilder	daf68dbbd2	Merge pull request #5701 from jonseymour/js-deduplicate-safety tsm: cache: improve thread safety of Cache.Deduplicate (see #5699)	2016-02-25 09:18:10 -07:00
Mark Rushakoff	e7bb855ab2	Merge pull request #5816 from influxdata/mr-database-stats Track stats for number of series, measurements	2016-02-25 08:13:04 -08:00
Ben Johnson	0dda9f6608	add remote execution This commit adds remote execution to the query engine.	2016-02-25 08:41:20 -07:00
Jon Seymour	4d98a1cf28	tsm: cache: remove unnecessary lock escalation. Previously, we needed a write lock on the cache because it was the only lock we had available to guard updates to entry.values and entry.needSort. However, now we have a entry-scoped lock for this purpose, we don't need the cache write lock for this purpose. Since merged() doesn't modify the .store or the c.snapshot.sort, there is no need for a write lock on the cache to protect the cache. So, we don't need to escalate here - we simply rely on the entry lock to protect the entries we are iterating over. Signed-off-by: Jon Seymour <jon@wildducktheories.com>	2016-02-26 01:31:54 +11:00
Edd Robinson	aa845cec7e	Check for shards needing conversion. Fixes #5723	2016-02-25 13:21:13 +00:00
Jason Wilder	452d77cbaf	tsm: cache: introduce entry locks. Based on @jwilder's alternative to the 'dirty' slice that featured in previous iterations of this fix. Suggested-by: Jason Wilder <jason@influxdb.com> Signed-off-by: Jon Seymour <jon@wildducktheories.com>	2016-02-26 00:05:38 +11:00
Jon Seymour	eb7eec078d	tsm: cache: introduce commit lock to Cache Currently two compactors can execute Engine.WriteSnapshot at once. This isn't thread safe since both threads want to make modifications to Cache.snapshot at the same time. This commit introduces a lock which is acquired during Snapshot() and released during ClearSnapshot(), ensuring that at most one thread executes within Engine.WriteSnapshot() at once. To ensure that we always release this lock, but only release the snapshot resources on a successful commit, we modify ClearSnapshot() to accept a boolean which indicates whether the write was successful or not and guarantee to call this function if Snapshot() has been called. Signed-off-by: Jon Seymour <jon@wildducktheories.com>	2016-02-25 12:10:37 +11:00
Jon Seymour	45d025db99	tsm: cache: add a tests to demonstrate thread safety vulnerabilities There are two tests that show two different one vulnerability. One test shows that Cache.Deduplicate modifies entries in a snapshot's store without a lock while cache readers are deduplicating those same entries while correctly locked. A second test shows that two threads trying to execute the methods that Engine.WriteSnapshot calls will cause concurrent, unsynchronized mutating access to the snapshot's store and entries. The tests fail at this commit and are fixed by subsequent commits. Signed-off-by: Jon Seymour <jon@wildducktheories.com>	2016-02-25 12:10:31 +11:00
Jon Seymour	d7d81f79da	tsm: cache: add a test that demonstrates concurrent reads are safe Signed-off-by: Jon Seymour <jon@wildducktheories.com>	2016-02-25 12:06:10 +11:00
Mark Rushakoff	fb83374389	Track stats for number of series, measurements Per database: track number of series and measurements Per measurement: track number of series	2016-02-24 08:10:16 -08:00
Edd Robinson	16995b6c23	Add ShardError to provide context about shard that errored	2016-02-24 13:33:07 +00:00
Jon Seymour	530b86ba7d	tsm: cache: restore the semantics of cachedBytes and memSize stats Fixes #5805. This commit undoes a regression introduced by #5789. Signed-off-by: Jon Seymour <jon@wildducktheories.com>	2016-02-24 06:16:46 +11:00
Jon Seymour	3475356dc9	tsm: cache: fix semantics of snapshotCount statistic to make it useful. Fix for #5804. The commit for #5789 rendered the semantics of snapshotCount statistic useless. This commit restores semantics that have diagnostic value to this statistic. Signed-off-by: Jon Seymour <jon@wildducktheories.com>	2016-02-24 06:13:54 +11:00
Jason Wilder	017c24c98e	Simplify cache snapshotting The Cache had support for taking multiple snapshots to support writing multiple snapshots to TSM files concurrently if that happened to be a bottleneck. In practice, this is never a bottleneck and we only run one snappshoting goroutine continously per shard which has worked well for all workloads. The multiple snapshot support introduces some unhandled failure scenarios where wal segments could be removed without writing them to TSM files. If a snapshot compaction fails to write due to transient disk errors, subsequent snapshots will continue, but the failed one will not be retried. When the subsequent ones succeeded, all closed wal segments are removed causing data loss. This change simplifies the snapshotting capability to ensure that there is only ever one snapshot. If one fails, the next snapshot will update the existing snapshot and retry all of old and new data. Fixes #5686	2016-02-23 09:38:51 -07:00
Jonathan A. Sternberg	50753de032	Merge pull request #5782 from influxdata/js-5777-audit-panics-in-influxql Remove the non-unreachable panics in the new query engine	2016-02-22 17:18:57 -05:00
Mark Rushakoff	191de2670c	Fix non-compiling test	2016-02-22 13:49:11 -08:00
Mark Rushakoff	fc5c8597ab	Merge pull request #5758 from influxdata/mr-disk-stats Track cache, WAL, filestore stats within tsm1 engine	2016-02-22 13:01:55 -08:00
Jason Wilder	aa2e878019	Fix cache not deduplicating points in some cases The cache had some incorrect logic for determine when a series needed to be deduplicated. The logic was checking for unsorted points and not considering duplicate points. This would manifest itself as many points (duplicate) points being returned from the cache and after a snapshot compaction run, the points would disappear because snapshot compaction always deduplicates and sorts the points. Added a test that reproduces the issue. Fixes #5719	2016-02-22 13:24:42 -07:00
Jonathan A. Sternberg	7a03df2af1	Remove the non-unreachable panics in the new query engine The only panics left are ones that should be unreachable unless there is a bug. Fixes #5777.	2016-02-22 12:52:43 -05:00
Jon Seymour	c93da21a61	tsm: cache: only use NewCache for engine cache's snapshots use a simpler constructor The intent of this change is to avoid writing caches created for snapshot cache instances into the tsm1_cache measurement. We can do this by avoiding use of the NewCache constructor. All other methods are only intended to be called from on the engine cache - never on a snapshot. Signed-off-by: Jon Seymour <jon@wildducktheories.com>	2016-02-22 15:17:43 +11:00
Jon Seymour	510ee2c790	tsm: cache: during writes, update the memSize statistic outside the lock Since we are not locking but relying on atomic arithmetic, use Add rather than Set. Will also result in slightly less garbage being created. Signed-off-by: Jon Seymour <jon@wildducktheories.com>	2016-02-22 08:26:35 +11:00
Jon Seymour	9c6efe99f1	tsm: cache: ensure all statistics are initialised on cache creation. The intent of this change is to ensure that all statistic fields of the resulting tsm1_cache measurement are initialized on initialization of the cache. That way, any consumer of those measurements doesn't have to deal with the null case. Signed-off-by: Jon Seymour <jon@wildducktheories.com>	2016-02-21 15:33:50 +11:00
Jon Seymour	6697c721fb	tsm: cache: add cache throughput related statistics. Complementing and extending the changes in #5758. Add 2 level statistics: * snapshotCount * cacheAgeMs Add 2 counter statistics * cachedBytes * WALCompactionTimeMs snapshotCount can be used to measure transient write errors that are causing snapshots to accumulate cacheAgeMs can be used to guage the level of write activity into the cache The differences between cachedBytes stats sampled at different times can be used to calculate cache throughput rates The ratio (cachedBytes-diskBytes)/WALCompactionTimeMs can be used calculate WAL compaction throughput. The ratio of difference between first and last WAL compaction time over the interval length is an estimate of percentage of cache throughput consumed. Signed-off-by: Jon Seymour <jon@wildducktheories.com>	2016-02-20 22:18:57 +11:00
Mark Rushakoff	602043e11b	Add disk stats for FileStore	2016-02-19 16:37:34 -08:00
Mark Rushakoff	d99c09cedd	Add stats for current and old WAL segment sizes	2016-02-19 16:37:34 -08:00
Mark Rushakoff	e76967efb6	Add stats to tsm1.Cache	2016-02-19 16:37:34 -08:00
Edd Robinson	99a7341701	Wire up DROP retention policy to TSDB store. Fixes #5653 and #5394. Previously dropping retention policies did not propogate to local TSDB shards. Instead, the retention policiess would just be removed from the Meta Store. This PR adds ensures that data associated with retention policies is removed, when the retention policy is dropped. Also, it cleans up a couple of other methods in `tsdb`, including the requirement to provide (redundant) shardIDs when deleting databases.	2016-02-19 11:15:00 +00:00
Joe LeGasse	dc8ed7953d	Remove custom binary-conversion functions Also cleaned up some excess allocations, and other cruft from the code	2016-02-18 13:56:35 -05:00
Ben Johnson	eb221a5adb	Merge pull request #5663 from benbjohnson/query-executor Refactor QueryExecutor (WIP)	2016-02-17 16:30:20 -07:00
Ben Johnson	e3b4b71c13	refactor query executor This commit moves the `QueryExecutor` to the `cluster` package and provides an interface to it inside the `influxql` package.	2016-02-17 15:13:56 -07:00
Ben Johnson	f7e04abef7	remove NaN from query engine This commit removes `math.NaN` returns from float iterators.	2016-02-17 14:11:31 -07:00
liang@qiniu.com	1ad0f933f4	Remove redundant wal files	2016-02-16 20:45:13 +08:00
Jon Seymour	ab702eb44a	doc: remove the implication that the wal directory is inside the shard directory. Signed-off-by: Jon Seymour <jon@wildducktheories.com>	2016-02-15 05:33:22 +11:00
Jon Seymour	ed0a112f8e	doc: Add an Errata section intended to capture clarifications prior to full revisions of the text. Signed-off-by: Jon Seymour <jon@wildducktheories.com>	2016-02-15 00:29:02 +11:00
Jon Seymour	5e563d53c1	doc: revise discussion about cache design The description of the cache design was out of date - reflecting an older design based on checkpoints and evictions. This revision updates the design to describe snapshots and also clarify that if compaction performance falls behind the inbound write rate then writes will fail. Updates based in part of clarifications provided by Jason Wilder. See https://goo.gl/L7AzVu Signed-off-by: Jon Seymour <jon@wildducktheories.com>	2016-02-15 00:29:02 +11:00
Jon Seymour	cdc7e28338	doc: rephrasing of how sets of SeriesIterators are generated. Signed-off-by: Jon Seymour <jon@wildducktheories.com>	2016-02-15 00:29:02 +11:00
Jon Seymour	58d1b7223a	doc: refine TSM file system layout description Minor improvements to phrasing to use the English word 'directory' and slight improvements to grammar.	2016-02-15 00:29:02 +11:00
Jon Seymour	285e0ad17a	doc: refine description of the conclusion of the compaction process. Signed-off-by: Jon Seymour <jon@wildducktheories.com>	2016-02-15 00:29:02 +11:00
Jon Seymour	008af05f7b	doc: various grammar/word-choice improvements in TSM design document Signed-off-by: Jon Seymour <jon@wildducktheories.com>	2016-02-15 00:29:02 +11:00
Jon Seymour	88598f78dc	doc: fix up some spelling errors/typos in .MD files Signed-off-by: Jon Seymour <jon@wildducktheories.com>	2016-02-15 00:29:02 +11:00
Jonathan A. Sternberg	559a11d0ab	Pass the implicit end time from the query executor to the select call The select call and the query executor would both calculate the time range, but in separate ways. The query executor needed some way to pass in the implicit end time that is placed there by the query executor. Fixes #5636.	2016-02-12 16:03:24 -05:00
Mark Rushakoff	fc9ab7a46f	Miscellaneous cleanup in tsdb package * When possible, initialize maps/slices to exact length/capacity * See slice benchmarks at https://gist.github.com/mark-rushakoff/b5650bd8f06bece0b9fd * Fixed some typos * Removed an unnecessary loop in stringset.intersect	2016-02-10 18:00:47 -08:00
Ben Johnson	0b3d367e5c	Merge pull request #5623 from influxdata/jw-query-panic Fix panic: runtime error: index out of range	2016-02-10 14:59:04 -07:00
Jason Wilder	0ce6dd1304	Fix panic: runtime error: index out of range There was a fix in 5b1791, but is not present in the current branch likely due to a rebase issue. The current code panics with a query like: select value from cpu group by host order by time desc limit 1 This fixes the panic as well as prevents #5193 from re-occurring. The issue is that agressively closing the cursors clears out the seeks slice so re-seeking will fail.	2016-02-10 14:00:58 -07:00
Justin Nuß	82c276756a	Lint tsdb and tsdb/engine package	2016-02-10 21:33:46 +01:00
Ben Johnson	d9a6a7340f	add canonical paths	2016-02-10 11:30:52 -07:00
Ben Johnson	5a0d1ab7c1	rename influxdb/influxdb to influxdata/influxdb This commit changes all the import and URL references from: github.com/influxdb/influxdb to: github.com/influxdata/influxdb	2016-02-10 10:26:18 -07:00
Jonathan A. Sternberg	2e7cf5328c	Fix go vet issues on 1.4 go 1.5 was being used to develop the query engine branch, but we aren't using 1.5 for master at the moment. This fixes issues that go vet brings up in 1.4 that don't exist in 1.5.	2016-02-10 09:40:30 -07:00
Jonathan A. Sternberg	d1f7c445e7	Modify iterators to work across shards Aux iterators now ask the iterator creator what series will be returned and determine which aux fields to create based on the results. The `tsdb.Shards` struct also creates a call iterator around the iterators returned from each shard.	2016-02-10 09:40:29 -07:00
Ben Johnson	32be4c250f	fix non-existent shard handling This commit removes `nil` shards returned from `tsdb.Store.Shards()` which caused panics in some SELECTs. This can occur if the meta store has created shards before the store or if the shards are distributed throughout a cluster. Fixes #5555	2016-02-10 09:40:29 -07:00
Jonathan A. Sternberg	c2d1206177	Implement the fill iterator Fill requires an additional function for IteratorCreator to retrieve the series that will be returned from the iterator. When fill is required for an aggregate, the IteratorCreator will be asked what series will be returned by the created iterator.	2016-02-10 09:40:29 -07:00
Ben Johnson	627cd9d486	add dedupe iterator	2016-02-10 09:40:29 -07:00
Ben Johnson	47c2bab74b	add SHOW TAG KEYS support	2016-02-10 09:40:28 -07:00
Ben Johnson	607750ab1b	add SHOW MEASUREMENTS iterator	2016-02-10 09:40:28 -07:00
Ben Johnson	2bdc9404ef	revert meta execution	2016-02-10 09:40:28 -07:00
Ben Johnson	6204350d65	fix math operations	2016-02-10 09:40:27 -07:00
Ben Johnson	b4cb770a7f	refactor aux iterators	2016-02-10 09:40:27 -07:00
Ben Johnson	b8918a780c	integer support	2016-02-10 09:40:25 -07:00
Jonathan A. Sternberg	583477064c	Check for `tsdb.EOF` when looking for the lowest timestamp of aux fields	2016-02-10 09:40:25 -07:00
Jonathan A. Sternberg	34f14424dd	Filter tags from the condition when building cursors on tsm1	2016-02-10 09:40:25 -07:00
Ben Johnson	00806de9b8	refactor query engine	2016-02-10 09:40:25 -07:00
Ben Johnson	57336bd6ee	fix conditionals	2016-02-10 09:40:24 -07:00
Ben Johnson	036382ee20	SLIMIT/SOFFSET	2016-02-10 09:40:24 -07:00
Ben Johnson	cde973f409	refactor query engine	2016-02-10 09:40:24 -07:00
Gabriel Levine	7d4217ab97	enabled golint for tsdb/engine/wal.go and wal_test.go and updated changelog.	2016-02-09 10:29:09 -05:00
Jason Wilder	2b3c640695	Fix reading too far in fileAccess.readBytes Fixes #5566	2016-02-08 09:08:57 -07:00
Jason Wilder	28ae8b6fe0	Merge pull request #5434 from runner-mei/tsm_tombstone_windows fix TSMReader.Delete() and all unit tests is pass in the windows	2016-02-04 16:27:26 -07:00
Jason Wilder	b635e516e5	Merge pull request #5485 from runner-mei/patch-7 fix munmap bug in the windows	2016-02-04 13:47:51 -07:00
Jason Wilder	5a124e0e0b	Merge pull request #5431 from runner-mei/patch-5 fix determine the file size	2016-02-04 10:24:05 -07:00
Edd Robinson	1bcb1d033f	Allow Close to be called multiple times safely	2016-02-03 10:20:22 +00:00
Edward Robinson	c7bbe6ef17	Remove engine on close	2016-02-03 10:19:42 +00:00
INADA Naoki	80a637904d	tsm1: Use unixnano instead of time.Time	2016-02-03 10:05:40 +09:00
INADA Naoki	771253256b	FloatValue uses unixnano instead of time.Time	2016-02-03 09:57:00 +09:00
INADA Naoki	898babf616	add float bench	2016-02-03 03:12:16 +09:00
Cory LaNou	a171089fce	fix vet error	2016-02-01 15:33:54 -06:00
Joe LeGasse	4b642fb365	Added tests for tsdb.greaterThan() Also updated style to match "Effective Go" recommendations.	2016-02-01 15:11:17 -05:00
Joe LeGasse	2bd85b3c9e	Update first() and last() to handle varying types	2016-02-01 11:38:08 -05:00
runner.mei	4ca47103b1	fix TSMReader.Delete() and all unit tests is pass in the windows	2016-01-31 11:32:08 +08:00
runner	bc992fea5e	fix munmap bug in the windows fix munmap bug in the windows fix munmap bug in the windows fix munmap bug in the windows fix munmap bug in the windows	2016-01-31 10:46:46 +08:00
runner	4b7fe70cd3	fix determine the file size fix determine the file size	2016-01-30 14:16:53 +08:00
runner.mei	53f7e03f72	fix TSMReader.Delete() and all unit tests is pass in the windows	2016-01-30 14:15:46 +08:00
Jason Wilder	924275b337	Fix panic preventing wal file truncation Fixes #5455	2016-01-28 21:50:51 -07:00
Jason Wilder	9528c3ea70	Merge pull request #5465 from influxdata/jw-remote-writes Optimize remote writes	2016-01-27 15:47:02 -07:00
Jason Wilder	1d165d38a9	Optimize Cache entry.add This reduces some of the lock contention when writing to the cache. When a new entry is created, it avoids an allocation. It also skips a check to see if we need to sorted if we already know it needs to sorted.	2016-01-27 14:26:42 -07:00
Ben Johnson	98baf078d0	tsm1 query performance improvements	2016-01-27 13:42:32 -07:00
Jason Wilder	372302bcbd	Reduce lock contention in Cache.WriteMulti A write-lock was taken the whole time, but we only need the write lock at the end.	2016-01-25 16:48:34 -07:00
Jason Wilder	5bee8880db	Reduce lock content in engine.WritePoints Writing the snapshot would deduplicate the snapshot points while still holding the engine write-lock. This can be expensive under high load and cause writes to back up and OOM the server. Instead, grab the snapshot under the lock and dedup it after releasing the lock. Possible fix for #5442	2016-01-25 15:37:34 -07:00
Jason Wilder	ad52d0fbd9	Fix tests	2016-01-21 15:30:09 -05:00
Paul Dix	f385945058	Update Server to work with new metaservice/client	2016-01-21 15:28:33 -05:00
Cory LaNou	9ec7a710c9	some misc refactoring on influxd startup	2016-01-21 15:28:32 -05:00
Cory LaNou	8d878fff91	buildable meta -> services/meta	2016-01-21 15:28:32 -05:00
Ben Johnson	ba7fc7d548	Merge pull request #5333 from benbjohnson/limit Limit raw query fetch	2016-01-12 17:06:41 -08:00
Jason Wilder	15d723dc77	Change default engine to tsm1 data engine config var is ignored now and you can only create tsm1 shards. Exists shards will work as is until they are migrated to tsm1 shards.	2016-01-11 12:02:36 -07:00
Ben Johnson	f5ee6a0713	limit raw query fetch This commit enforces a limit on `RawMapper` so that it will not produce more values than are specified by the LIMIT clause. Previously the mapper would read up to the chunk size and the values would be limited afterward.	2016-01-11 09:01:49 -07:00
Jason Wilder	24f1bcfd20	Remove Dev prefix from tsm engine/tx	2016-01-10 16:43:36 -07:00
Jason Wilder	5b179113fc	Don't close tsm cursor prematurely We were closing the cursor when we read the last block which caused the internal state to be cleared. In a group by query, we seeked multiple times so depending on the group by interval and how the data was laid out in the blocks, we woudl close the cursor and the last block would get skipped. Fixes #5193	2016-01-10 15:26:01 -07:00
Jason Wilder	3c45015311	Remove MAP_POPULATE This may be causing slow restart times for systems with many large TSM files. What I believe is happening at startup in these cases is that multiple goroutines are started to load each TSM file concurrently. The kernel appears to serialize mmap calls from the same process so all of the goroutines end up getting blocked on the actual mmap system call. MAP_POPULATE instruct the kernel to pre-fault the page table for the files and triggers read-ahead of the pages. For larger, 2GB files, this makes the mmap call more expensive and slower. When there are many of these files and calls it is possible to fill all available memory with pagecache. In this case, the OS will end up pre-faulting pages from one file and have to remove pages that it just loaded from another files causing slowness. MAP_POPULATE may also be cause much more data to be pre-faulted than necessary. To load a file, we just need to scan the index at the end of the file. MAP_POPULATE is likely causing the whole file to be loaded when it won't actually be accessed for a while (or at all). Might fix issue #5311.	2016-01-08 08:45:27 -07:00
Jason Wilder	756421ec4a	Look for fully compacted block in addition to max size during compaction Some data shapes would cause files to grow larger than the max size more quickly which resulted in them getting skipped by the full compaction planner at times. Some datasets that could make this happen are very large keys or very large numbers of keys (10M). When this happened, multiple max sized files would accumulate but the blocks would not be full. When the shard went cold for writes, these files would get recompacted down to the optimal size, but a lot of space would be wasted in the mean time.	2016-01-07 15:18:42 -07:00
Jason Wilder	faf8ee17fa	Fix typo	2016-01-06 12:53:04 -07:00
Jason Wilder	d2b7c03175	Re-use the series key Avoid allocating the string twice.	2016-01-06 12:52:13 -07:00
Jason Wilder	2f7a0090c1	Don't allocate a pre-sized buffer for each cursor This is contributing to some of the high memory usage on queries and possibly some OOMs. This is slightly slower, but removing it allows some fairly large count queries over 5M series to complete instead of crashing the process using tsm1 engine.	2016-01-06 10:50:38 -07:00
Jason Wilder	6f577cfef5	Reduce allocations when compacting Key() returned the key and the entries. We did not always need the entries so they would be allocated and ignored. Added a KeyAt func that just returns the key to avoid the unnecesary entries allocation.	2016-01-05 16:16:44 -07:00
Jason Wilder	9a9ccab560	Reduce allocation in wal encoder Use sync.Pool for some temporary buffers used while encoding instead of allocatin new ones each time. Also increased the default buffer size which might be too small. Probably need to make this a config var.	2016-01-05 16:12:25 -07:00
Jason Wilder	ee54a1e791	Write TSM data directly to writer We were buffering up the data to write into byte slices to reduce IO calls but at larger sizes, this causes memory to spike. The TSMWriter was switched to use a bufio.Writer internally so this byte slice buffering is unnecessary and costly now.	2016-01-05 14:46:07 -07:00
Jason Wilder	d2889ecd6a	Avoid creating slices of all keys during compaction	2016-01-05 09:38:00 -07:00
Jason Wilder	7794b9c5d4	Fix panic: runtime error: slice bounds out of range The block count was an uint16 when incrementing the index location which was an int32. This caused the value the uint16 value to overflow before the index location was incremented causing the wrong location to be read on the next iteration of the loop. This triggers the slice out of range errors. Added a test that recreates the panic seen in #5257 and possibly #5202 which is older code. Fixes #5257	2016-01-04 11:20:24 -07:00
Paul Dix	49d480cb0c	Fix races in backup/restore	2015-12-31 08:42:01 -05:00
Paul Dix	5974d37649	Fix backup test to mock out compaction	2015-12-31 08:15:13 -05:00
Paul Dix	9cede5fb71	Address PR comments	2015-12-30 18:06:51 -05:00
Paul Dix	26e1c6464a	Update backup to address PR comments	2015-12-30 18:06:51 -05:00
Paul Dix	59fbd371fc	Implement backup/restore for TSM. This changes backup and restore to work for TSM. It breaks it for b1 and bz1, but since those are getting removed it's ok. The backup runs against any host that is specified and can backup either the metasstore, a database, specific retention policy, or a specific shard. It can also take incremental backups with the `since` flag, which will only backup TSM files that have been created since that timestamp. The backup is safe to run online. However, for shards that are still hot for writes, they won't be able to create new TSM files while the backup for that single shard runs. If the backup isn't too large and the write throughput isn't too high this shouldn't be a problem since the writes will just go into the WAL cache.	2015-12-30 18:06:50 -05:00
Jason Wilder	b6da176a4b	Fix direct index size not calculated	2015-12-23 18:01:11 -07:00
Jason Wilder	f9ae8077da	Allow compactions to run when files have tombstones	2015-12-23 18:01:11 -07:00
Jason Wilder	a38c95ec85	Update compactions to run concurrently This has a few changes in it (unfortuantely). The main change is to run compactions concurrently. While implementing this, a few query and performance bugs showed up that are also fixed by this commit.	2015-12-23 18:01:11 -07:00
Jason Wilder	48d4156eac	Fix blocks not sorted correctly when chunking	2015-12-23 18:01:11 -07:00
Jason Wilder	bb2562b2ab	Return CompactionGroups from planning	2015-12-23 18:01:11 -07:00
Jason Wilder	d0ec0a15e2	Fix wrong test data setup	2015-12-23 18:01:11 -07:00
Ady	5c888b3673	Merge branch 'master' of https://github.com/influxdb/influxdb into mvadu-patch-4358 Trying to get to latest master from influxdb	2015-12-19 01:45:07 +05:30
Jason Wilder	7e97b0eafd	Fix rename temp file on windows	2015-12-18 11:57:37 -07:00
Jason Wilder	611017f4ed	Add comments	2015-12-18 10:00:07 -07:00
Jason Wilder	930174bf4d	Handle calling WriteBlock with no data gracefully	2015-12-18 09:57:16 -07:00
Jason Wilder	6bc7765b88	Handle calling write with no values to TSMWriter gracefully	2015-12-18 09:52:53 -07:00
Jason Wilder	421a127f11	Add indirectIndex.UnmarshalBinary benchmark	2015-12-17 15:38:51 -07:00
Jason Wilder	8c7e11f4cf	Aggressively clean up KeyCursor resources	2015-12-17 12:51:51 -07:00
Jason Wilder	fd2a409ea3	Skip decoding blocks that are already full	2015-12-17 12:47:05 -07:00
Jason Wilder	825296ddd8	Add comments	2015-12-16 11:30:06 -07:00
Jason Wilder	88324bf61c	Optimize indirectIndex.UnmarshalBinary further	2015-12-16 11:28:13 -07:00
Jason Wilder	70d1f45058	Load TSM files concurrently	2015-12-16 11:28:12 -07:00
Jason Wilder	737871268b	Speed up indirectIndex.UnmarshalBinary Remove a bunch of unnecessary allocations to improve startup times.	2015-12-16 11:16:17 -07:00
Jason Wilder	3893bc60e1	Speed up TSM compactor Just keep the current block for each iterator in the buffers.	2015-12-16 11:16:17 -07:00
Jason Wilder	00f570441b	Convert TSMKeyIterator to return blocks	2015-12-16 11:16:17 -07:00
Jason Wilder	59a57d8f73	Convert CacheKeyIterator to return encoded blocks	2015-12-16 11:16:17 -07:00
Jason Wilder	0623648140	Add chunking support back to TSMKeyIterator Was removed when MergeIterator was deleted.	2015-12-16 11:16:17 -07:00
Jason Wilder	31b97c3fe0	Add max points per block back for CacheKeyIterator Was removed when MergeIterator was removeed.	2015-12-16 11:16:16 -07:00
Jason Wilder	45e87cdfe4	Strip checksum when returning block from ReadBytes	2015-12-16 11:16:16 -07:00
Jason Wilder	97435b9124	Return minTime/maxTime from BlockIterator.Read	2015-12-16 11:16:16 -07:00
Jason Wilder	ce6de9728e	Add test for BlockIterator with multiple blocks for a key	2015-12-16 11:16:16 -07:00
Jason Wilder	4a3037814f	Add WriteBlock to TSMWriter	2015-12-16 11:16:16 -07:00
Jason Wilder	d99c1f944e	Add BlockIterator for reading TSM blocks without decoding	2015-12-16 11:16:16 -07:00
Jason Wilder	928aef04cd	Split data_file.go into reader.go and writer.go	2015-12-16 11:16:16 -07:00
Philip O'Toole	47317d73b4	Merge pull request #5131 from influxdb/site-fixes Default data logging to on	2015-12-16 10:03:25 -08:00
Alexandre Viau	ad1044dde9	typo: unkown -> unknown	2015-12-15 18:10:47 -05:00
Philip O'Toole	d45048455a	Default data logging to on	2015-12-15 13:15:38 -08:00
Philip O'Toole	0e4bc275d8	Merge pull request #5115 from influxdb/site-fixes Log TSM initialization	2015-12-15 13:13:17 -08:00
Philip O'Toole	01ac0b3f23	Tweak compaction log messages	2015-12-15 10:33:13 -08:00
dgnorton	d89e233567	Merge pull request #5100 from influxdb/dgn-fix-4303 fix #4303: don't drop from multiple databases	2015-12-15 07:27:05 -05:00
Philip O'Toole	a6cdb5229d	Log tsm initialization	2015-12-14 15:50:56 -08:00
David Norton	3014fb90e4	fix #4303 : don't drop from multiple databases	2015-12-12 13:54:23 -05:00
Philip O'Toole	75764517f6	Merge pull request #5082 from li-ang/fix_x Fix wrong value of countCompacting in WAL	2015-12-11 10:07:56 -08:00
Philip O'Toole	03f8cd3956	Add comment explaining magic number	2015-12-10 11:46:40 -08:00
Jason Wilder	631ecc23de	Fix growing destination buffer during WAL entry encoding The test to see if the destination buffer for encoding and decoding a WAL entry was broken and would cause a panic if there were large batches that would overflow the buffer size. Fixes #5075	2015-12-10 11:46:40 -08:00
liang@qiniu.com	34bdffdb00	Fix wrong value of countCompacting in wal	2015-12-10 17:47:20 +08:00
Nathaniel Cook	b7000c80dd	count with fill(none) will drop 0 valued intervals	2015-12-09 15:20:47 -07:00
Ady	07c0939fe1	Added logic To let the memeory mapped files to renamed by OS. Now a copy is created in memory with SHARED_DELETE flag, so that OS is free to rename or delete original file	2015-12-10 01:07:50 +05:30
Philip O'Toole	da08304780	Merge pull request #4940 from li-ang/fix_aggregative_query_err Fix distributed aggregative query error	2015-12-09 11:04:40 -08:00
Jason Wilder	992aea7bd3	Merge pull request #5060 from influxdb/jw-drop-db Cancel writing TSM files when engine closes	2015-12-08 16:16:07 -07:00
Paul Dix	b192136887	Merge pull request #5058 from influxdb/pd-update-compaction-logic Update TSM compaction logic	2015-12-08 18:14:15 -05:00
Paul Dix	27cc2ea0cc	Update compact.Plan	2015-12-08 18:01:31 -05:00
Jason Wilder	d7cff651d1	Cancel writing TSM files when engine closes If the engine is closed while a compaction is going on, the close call blocks until the goroutine exits. This could be several minutes because the control does not return back up to the channel selector while there is still data to write.	2015-12-08 15:41:53 -07:00

... 2 3 4 5 6 ...

1126 Commits (2cbddb3efd851d408b9eb0676e168b27dbe7eb79)