influxdb

Commit Graph

Author	SHA1	Message	Date
Raphael Taylor-Davies	746b18bbed	feat: allow overriding CLI request timeout (#3824 ) * feat: allow overriding CLI request timeout * chore: rename to --rpc-timeout to avoid name collision Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-02-23 11:35:45 +00:00
dependabot[bot]	b63f920d4c	chore(deps): Bump parquet from 9.0.2 to 9.1.0 (#3828 ) * chore(deps): Bump parquet from 9.0.2 to 9.1.0 Bumps [parquet](https://github.com/apache/arrow-rs) from 9.0.2 to 9.1.0. - [Release notes](https://github.com/apache/arrow-rs/releases) - [Changelog](https://github.com/apache/arrow-rs/blob/master/CHANGELOG.md) - [Commits](https://github.com/apache/arrow-rs/compare/9.0.2...9.1.0) --- updated-dependencies: - dependency-name: parquet dependency-type: direct:production update-type: version-update:semver-minor ... Signed-off-by: dependabot[bot] <support@github.com> * chore: update chunk size test Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com> Co-authored-by: Raphael Taylor-Davies <r.taylordavies@googlemail.com> Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-02-23 11:25:15 +00:00
dependabot[bot]	5a79b3a68b	chore(deps): Bump arrow-flight from 9.0.2 to 9.1.0 (#3829 ) Bumps [arrow-flight](https://github.com/apache/arrow-rs) from 9.0.2 to 9.1.0. - [Release notes](https://github.com/apache/arrow-rs/releases) - [Changelog](https://github.com/apache/arrow-rs/blob/master/CHANGELOG.md) - [Commits](https://github.com/apache/arrow-rs/compare/9.0.2...9.1.0) --- updated-dependencies: - dependency-name: arrow-flight dependency-type: direct:production update-type: version-update:semver-minor ... Signed-off-by: dependabot[bot] <support@github.com> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com> Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-02-23 11:03:22 +00:00
dependabot[bot]	3b7d31c88a	chore(deps): Bump arrow from 9.0.2 to 9.1.0 (#3826 ) Bumps [arrow](https://github.com/apache/arrow-rs) from 9.0.2 to 9.1.0. - [Release notes](https://github.com/apache/arrow-rs/releases) - [Changelog](https://github.com/apache/arrow-rs/blob/master/CHANGELOG.md) - [Commits](https://github.com/apache/arrow-rs/compare/9.0.2...9.1.0) --- updated-dependencies: - dependency-name: arrow dependency-type: direct:production update-type: version-update:semver-minor ... Signed-off-by: dependabot[bot] <support@github.com> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2022-02-23 09:25:46 +00:00
dependabot[bot]	ad3868ed7c	chore(deps): Bump tokio from 1.16.1 to 1.17.0 (#3814 ) * chore(deps): Bump tokio from 1.16.1 to 1.17.0 Bumps [tokio](https://github.com/tokio-rs/tokio) from 1.16.1 to 1.17.0. - [Release notes](https://github.com/tokio-rs/tokio/releases) - [Commits](https://github.com/tokio-rs/tokio/compare/tokio-1.16.1...tokio-1.17.0) --- updated-dependencies: - dependency-name: tokio dependency-type: direct:production update-type: version-update:semver-minor ... Signed-off-by: dependabot[bot] <support@github.com> * build: update workspace-hack Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com> Co-authored-by: Dom Dwyer <dom@itsallbroken.com> Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-02-22 16:27:43 +00:00
dependabot[bot]	65ab5213e5	chore(deps): Bump clap from 3.0.14 to 3.1.1 (#3809 ) Bumps [clap](https://github.com/clap-rs/clap) from 3.0.14 to 3.1.1. - [Release notes](https://github.com/clap-rs/clap/releases) - [Changelog](https://github.com/clap-rs/clap/blob/master/CHANGELOG.md) - [Commits](https://github.com/clap-rs/clap/compare/v3.0.14...v3.1.1) --- updated-dependencies: - dependency-name: clap dependency-type: direct:production update-type: version-update:semver-minor ... Signed-off-by: dependabot[bot] <support@github.com> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2022-02-22 14:51:53 +00:00
Raphael Taylor-Davies	1960645055	feat: add wildcard support to persist partition CLI command (#3790 ) * feat: add wildcard support to persist partition * chore: fmt Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-02-22 14:21:06 +00:00
Raphael Taylor-Davies	0229147909	feat: preserve catalog on error (#1522 ) (#3802 ) Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-02-22 14:10:20 +00:00
kodiakhq[bot]	326871f149	Merge branch 'main' into dom/router-handler-chain	2022-02-22 13:49:09 +00:00
Luke Bond	0f012de70c	feat: adding compactor CLI command and crate Closes: #3777	2022-02-21 12:24:09 +00:00
Raphael Taylor-Davies	39c42678d7	feat: trigger persistence if over soft limit and no evictable chunks (#3791 ) * feat: trigger persistence if over soft limit and no evictable chunks * chore: fmt * fix: avoid test_full_lifecycle exceeding soft limit * fix: don't expect chunk to be unloaded * feat: only trigger if no outstanding persist Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-02-21 10:14:27 +00:00
Dom Dwyer	bb132b61ad	refactor: chain DML handlers The router is composed of several DML handlers called in sequence in order to construct the full request handling pipeline. Prior to this commit, each handler nested the next handler it calls internally, producing a nested call chain that resulted metrics (added in #3764) recording cumulative latency like this: ┌ ─ │ ┌───────────────┐ │ NS Creation │ │ └───────────────┘ │ ┌───────────────┐ │ │ │ Partitioner │ │ └───────────────┘ │ │ │ │ │ Cumulative │ │ │ ┌───────────────┐ Timings 1.5s 1s │ etc... │ │ │ │ └───────────────┘ │ │ │ │ │ │ ┌───────────────┐ │ │ │ Partitioner │ │ └───────────────┘ │ ┌───────────────┐ │ NS Creation │ │ └───────────────┘ └ ─ This meant it was hard to determine the latency of a single handler without knowing (and subtracting the latency of) all the child handlers it calls. This commit replaces the intrusive nested handler call chain with an external Chain combinator type to compose together individual handlers, resulting in correct per-handler timings and simpler code/tests: ┌───────────────┐ │ NS Creation │ └───────────────┘ │ .5s ┌───────────────┐ └───────▶│ Partitioner │ └───────────────┘ │ 1s ┌───────────────┐ └───▶│ etc... │ └───────────────┘	2022-02-18 14:19:53 +00:00
Raphael Taylor-Davies	83cba3d2fb	feat: template static router config (#3781 ) * feat: template static router config * chore: lint and improved failure output * chore: clarify docs Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-02-18 10:53:10 +00:00
Marco Neumann	f54ef92b77	fix: supervise and shutdown ingester background tasks (#3769 ) * fix: supervise and shutdown ingester background tasks Closes #3761. Closes #3762. * docs: improve wording Co-authored-by: Raphael Taylor-Davies <1781103+tustvold@users.noreply.github.com> * test: join/shutdown handling for ingester Co-authored-by: Raphael Taylor-Davies <1781103+tustvold@users.noreply.github.com>	2022-02-18 09:35:29 +00:00
Andrew Lamb	9588b43a90	fix: Make errors in rewriting return `Error` rather than a `panic` (#3767 ) * test: add test for predicate errors * fix: Return errors properly rather than panic * fix: handle errors in influxrpc planner * fix: appease clippy * fix: tests Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-02-17 15:39:14 +00:00
kodiakhq[bot]	c89fa3701e	Merge branch 'main' into dom/router2-metrics	2022-02-17 15:16:39 +00:00
Edd Robinson	7a2b43f1fb	refactor: emit trace ID information	2022-02-16 19:01:14 +00:00
Edd Robinson	e4e9b56930	feat: add support for auto-generating a query trace	2022-02-16 18:08:49 +00:00
Dom Dwyer	d6d0ae8d80	feat: add instrumentation to request pipeline Wraps the sharded write buffer, schema validator, partitioner and overall request handler in instrumentation to record call latencies and export them via the /metrics endpoint.	2022-02-16 14:00:49 +00:00
Dom Dwyer	92fe507e52	feat: instrumented namespace cache Decorates the NamespaceCache with a set of cache get hit/miss counters, and put insert/update counters to expose cache behaviour.	2022-02-16 14:00:49 +00:00
kodiakhq[bot]	d0965bb0b2	Merge branch 'main' into dom/mb-partitioning	2022-02-16 11:30:42 +00:00
Paul Dix	f542045485	feat: wire up persistence in ingester (#3685 ) This adds persistence into the ingester with a lifecycle manager. The persist operation must still be updated to keep track of the min_unpersisted_sequence_number for each sequencer.	2022-02-16 00:13:40 +00:00
Edd Robinson	7ac9e216c4	refactor: use same log message	2022-02-15 14:36:55 +00:00
Edd Robinson	8a5ea29190	refactor: add measurement to log	2022-02-15 14:31:26 +00:00
Marco Neumann	44ee0166a0	fix: start Kafka write buffer stream at "earliest" offset, not at "0" (#3748 )	2022-02-15 13:36:59 +00:00
Marco Neumann	9e7a27b344	fix: default Kafka topic name is `iox-shared` (#3747 ) Do NOT use underscores in the Kafka topic because this is not supported by Kafka. This was initially fixed by #3555 but reverted by #3623.	2022-02-15 12:34:46 +00:00
Andrew Lamb	a30803e692	chore: Update datafusion, update `arrow`/`parquet`/`arrow-flight` to 9.0 (#3733 ) * chore: Update datafusion * chore: Update arrow * fix: missing updates * chore: Update cargo.lock * fix: update for smaller parquet size * fix: update test for smaller parquet files * test: ensure parquet_file tests write multiple row groups * fix: update callsite * fix: Update for tests * fix: harkari * fix: use IoxObjectStore::existing Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-02-15 12:10:24 +00:00
Dom Dwyer	e055800039	refactor: enable Partitioner in request pipeline Adds the Partitioner DML handler into the handler stack, modifying the input types of down-stream handlers to accept the partitioned data.	2022-02-15 11:34:33 +00:00
dependabot[bot]	89105ccfab	chore(deps): Bump tokio-util from 0.6.9 to 0.7.0 (#3743 ) Bumps [tokio-util](https://github.com/tokio-rs/tokio) from 0.6.9 to 0.7.0. - [Release notes](https://github.com/tokio-rs/tokio/releases) - [Commits](https://github.com/tokio-rs/tokio/commits) --- updated-dependencies: - dependency-name: tokio-util dependency-type: direct:production update-type: version-update:semver-minor ... Signed-off-by: dependabot[bot] <support@github.com> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2022-02-15 11:33:41 +00:00
Dom Dwyer	e99922d518	refactor: parametrise DML handler input type Allow a DML handler to specify the write input type on which it operates. This allows us to construct a write handler pipeline that transforms the request as it passes through the various handlers. We'll use this to implement a handler that annotates a normal set of table writes with the partition key, modifying downstream handlers to expect this annotated input.	2022-02-15 11:23:45 +00:00
Marco Neumann	c6e374a025	feat: allow catalog access w/o a transaction (#3735 ) * feat: allow catalog access w/o a transaction Now the caller has the full control if they want to use a transaction or not. * fix: remove non-transaction-safe `create_many` * fix: remove unnecessary transactions	2022-02-15 10:15:36 +00:00
dependabot[bot]	60a7f87645	chore(deps): Bump serde_json from 1.0.78 to 1.0.79 (#3739 ) Bumps [serde_json](https://github.com/serde-rs/json) from 1.0.78 to 1.0.79. - [Release notes](https://github.com/serde-rs/json/releases) - [Commits](https://github.com/serde-rs/json/compare/v1.0.78...v1.0.79) --- updated-dependencies: - dependency-name: serde_json dependency-type: direct:production update-type: version-update:semver-patch ... Signed-off-by: dependabot[bot] <support@github.com> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com> Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-02-14 20:42:54 +00:00
Raphael Taylor-Davies	26fd5273f0	feat: static database configuration (#2436 ) (#3732 ) * feat: static database configuration (#2436) * chore: fmt * feat: don't base64 encode UUIDs in ServerConfigFile Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-02-14 19:42:49 +00:00
Raphael Taylor-Davies	c79050254f	refactor: traitify database configuration (#2436 ) (#3730 ) * refactor: traitify database configuration (#2436) * chore: review feedback Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-02-13 09:26:44 +00:00
Raphael Taylor-Davies	866777ecd2	feat: static router configuration (#2436 ) (#3725 ) Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-02-11 14:09:37 +00:00
Raphael Taylor-Davies	4e3f66ed07	feat: CLI and gRPC APIs for shutting down and restarting databases (#3720 ) * feat: allow catalog wipe and rebuild whilst shutdown * feat: CLI and gRPC APIs for shutting down and restarting databases * feat: add ability to skip replay on restart * fix: test_wipe_persisted_catalog_error_db_exists * fix: wipe_preserved_catalog	2022-02-11 10:14:43 +00:00
Raphael Taylor-Davies	910f381355	refactor: require UUID to create Database (#3715 ) * refactor: require UUID to create Database * chore: review feedback * chore: fmt Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-02-10 20:04:06 +00:00
Raphael Taylor-Davies	b1190262b7	feat: restartable `Database` (#3368 ) (#3711 ) * feat: restartable `Database` (#3368) * chore: fmt * fix: wipe_preserved_catalog * chore: review feedback Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-02-10 18:32:05 +00:00
Andrew Lamb	d9f331ba2a	chore: update datafusion, stop repartitioning so aggressively (#3633 ) * chore: update datafusion * fix: Update to use new datafusion api * chore: update expected plans * fix: support zero output partitions * fix: update test * fix: Update for new DataFusion API * fix: newly added system table * fix: update cargo lock	2022-02-09 19:53:41 +00:00
Carol (Nichols \|\| Goulding)	73828323ac	feat: Ingester Flight gRPC API (#3623 ) * feat: Add a way to run ingester with an in-memory catalog from the CLI If you set the --catalog-dsn string to "mem", rather than using that as a Postgres connection URL, create an in-memory catalog. Planning on using this in tests, so not documenting. * fix: Set default topic to the same value as SHARED_KAFKA_TOPIC Namely, both should use an underscore. I don't think there's a way to directly share these values between a constant and an annotation. * feat: Add a flight API (handshake only) to ingester * fix: Create partitions if using file-based write buffer * fix: Change the server fixture to handle ingester server type For now, the ingester doesn't implement the deployment API. Not sure if it should or not. * feat: Start implementing ingester do_get, namely decoding the query Skip serialization of the predicate for the moment. * refactor: Rename ingest protos to ingester to match crate name * refactor: Rename QueryResults to QueryData * feat: Move ingester flight client to new querier crate * fix: Off by one error, different starting indexes in sequencers * fix: Create new CLI argument to pick the catalog type * fix: Create a CLI option to set the number of topics to auto-create in the write buffer * fix: Check the arrow flight service's health to tell that the ingester gRPC is up * fix: Set postgres as the default catalog type * fix: Return an error rather than panicking if CLI args aren't right	2022-02-09 19:07:44 +00:00
Edd Robinson	2334e779eb	feat: implement read_window_aggregate sub-command	2022-02-09 12:32:48 +00:00
Edd Robinson	0774e1d328	feat: add read_window_aggregate request builder	2022-02-09 12:32:48 +00:00
Marco Neumann	4bddab56e2	feat: create new sequencers in ingester on demand (#3671 ) There is no need to introduce yet another admin action to do that. If the sequencer does not exist yet, we can just create it and set the `min_unpersisted_sequence_number` to 0 (which is done be `create_or_get`).	2022-02-09 12:26:30 +00:00
Edd Robinson	dfa6fd8579	feat: add quiet option to storage	2022-02-08 21:27:29 +00:00
Edd Robinson	11855a5eff	feat: add format flag	2022-02-08 21:15:07 +00:00
Edd Robinson	c175ccd1b4	feat: make stop/stop/predicate global (#3681 )	2022-02-08 20:06:47 +00:00
kodiakhq[bot]	ace76cef14	Merge branch 'main' into dom/sharded-cache	2022-02-08 16:09:48 +00:00
Paul Dix	59b2141c0b	feat: Add lifecycle manager to ingester (#3645 ) This adds the lifecycle manager to the ingester. It will trigger based on a threshold for max partition size or age or based on keeping total memory under a certain threshold. It defines a new interface for a persister, which is stubbed out for IngesterData. I'm not sure yet how persistence errors should be handled. The assumption here is that the persister continues to retry persistence forever until it succeeds. There is one scenario I can think of that may cause this lifecycle manager problems. If a single partition is very high throughput, it could cause things to back up as persistence is not parallelized within a single partition. Any given partition can currently only run one persistence operation at a time. We can address this later. Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-02-08 15:23:40 +00:00
Marco Neumann	5de4d6203f	refactor: catalog transaction (#3660 ) * refactor: catalog Unit of Work (= transaction) Setup an inteface to handle Units of Work within our catalog. Previously both the Postgres and the in-mem backend used "mini-transactions on demand". Now the caller has a clear way to establish boundaries and gets read and write isolation. A single `Arc<dyn Catalog>` can create as many `Box<dyn UnitOfWork>` as you like, but note that depending on the backend you may not scale infinitely (postgres will likely impose certain limits and the in-mem backend limits concurrency to 1 to keep things simple). * docs: improve wording Co-authored-by: Andrew Lamb <andrew@nerdnetworks.org> * refactor: rename Unit of Work to Transaction * test: improve `test_txn_isolation` * feat: clearify transaction drop semantics Co-authored-by: Andrew Lamb <andrew@nerdnetworks.org> Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-02-08 13:38:33 +00:00
kodiakhq[bot]	4567800901	Merge branch 'main' into er/feat/tag_values_cli	2022-02-08 13:07:59 +00:00
Raphael Taylor-Davies	be662ec731	feat: lazy query log! (#3654 ) * feat: lazy query log * chore: fmt * chore: review feedback Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-02-08 13:07:28 +00:00
Edd Robinson	6c10e1e901	feat: support _measurement/_field tag keys	2022-02-08 11:32:28 +00:00
Edd Robinson	eb733042ca	feat: add support for tag_values cli	2022-02-07 22:02:29 +00:00
Edd Robinson	38a889ecf6	refactor: remove unnecessary struct	2022-02-07 22:02:29 +00:00
Marco Neumann	d9cc9f5a2a	feat: expose write buffer connection config via CLI (#3651 ) * feat: improve rskafka config error messages * feat: expose write buffer connection config via CLI	2022-02-07 16:24:28 +00:00
Marco Neumann	977ccc1989	fix: use a single metric registry for ingester (#3652 ) With this change write buffer ingestion metrics are showing up under `/metrics` Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-02-07 15:56:54 +00:00
Edd Robinson	87ac926e06	feat: add queries system table (#3655 ) Co-authored-by: Andrew Lamb <andrew@nerdnetworks.org>	2022-02-07 15:26:06 +00:00
Carol (Nichols \|\| Goulding)	2e30483f1f	refactor: Remove predicate module from predicate crate (#3648 ) Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-02-07 14:54:07 +00:00
Marco Neumann	e2db1df11f	refactor: improve writer buffer consumer interface (#3631 ) * refactor: improve writer buffer consumer interface The change looks huge but is actually rather simple. To understand the interface change, let me first explain what we want: - be able to fetch watermarks for any sequencer - have streams: - each streams tracks a sequencer and has an offset state (no read multiplexing) - we can seek a stream - seeking and streaming cannot be done at the same time (that would be weird and likely leads to many bugs both in write buffer and in the user code) - ideally we don't need to create streams of all sequencers but can choose a subset Before this change we had one mutable consumer struct where you can get all streams and watermark functions (this mutable-borrows the consumer) or you can seek a single stream (this also mutable-borrows the consumer). This is a bit weird for multiple reasons: - you cannot seek a single stream without dropping all of them - the mutable-borrow construct makes it really difficult to pass the streams into separate threads - the consumer is boxed (because its mutable) which makes it more difficult to handle in a large-scale application What this change does is the following: - you have an immutable consumer (similar to the producer) - the consumer offers the following methods: - get the set of sequencer IDs - get watermark for any sequencer - get a stream handler (see next point) for any sequencer - the stream handler captures the stream state (offset) and provides you a standard `Stream<_>` interface as well as a seek function. Mutable-borrows ensure that you cannot use both at the same time. The stream handler provides you the stream via `handler.stream()`. It doesn't implement `Stream<_>` itself because the way boxing, dynamic dispatch work, and pinning interact (i.e. I couldn't get it to work without the indirection). As a bonus point (which we don't use however) you can now create multiple streams for the same sequencer and they all have their own offset. * fix: review comments Co-authored-by: Carol (Nichols \|\| Goulding) <193874+carols10cents@users.noreply.github.com> Co-authored-by: Carol (Nichols \|\| Goulding) <193874+carols10cents@users.noreply.github.com> Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-02-07 12:24:17 +00:00
Edd Robinson	a52c0a26e6	feat: print read filter results	2022-02-04 22:14:22 +00:00
Edd Robinson	d328b37803	feat: teach IOx to convert RPC frames into Recordbatches	2022-02-04 18:34:54 +00:00
Edd Robinson	4cdaaf96bf	refactor: clean up errors	2022-02-04 18:34:54 +00:00
Edd Robinson	ea0ece8b4b	feat: issue read_filter request	2022-02-04 18:34:54 +00:00
Dom Dwyer	0b044b95fb	perf: use sharded namespace cache Enables the ShardedCache for the namespace schema cache.	2022-02-04 16:12:51 +00:00
Dom Dwyer	026a557c0b	refactor: rename TableNamespaceSharder Rename to JumpHash and expose the hashing internals for reuse (outside of only table & namespace sharding).	2022-02-04 15:56:09 +00:00
Dom Dwyer	0fd122e365	refactor: "inf" retention const Adds the iox_catalog::INFINITE_RETENTION_POLICY constant.	2022-02-04 15:35:33 +00:00
Dom Dwyer	f1ba50f40b	feat: resolve query pool ID at startup This commit adds a --query-pool flag to router2, used to upsert a catalog record at startup. Auto-created namespaces will reference this query pool. This is for testing only and will be removed in a future commit.	2022-02-04 15:35:30 +00:00
Dom Dwyer	aefc70a9ea	feat(router2): namespace auto-creation Decorate the existing request handler pipeline with a layer that implicitly creates the namespace when a write request is received.	2022-02-04 15:34:15 +00:00
Marco Neumann	0c01044677	fix: partition range in ingester CLI has INCLUSIVE end (#3641 )	2022-02-04 13:41:57 +00:00
Marco Neumann	d2ccf23263	fix: use standard DSN argument for router2 CLI (#3632 ) - support long-form (instead of relying on positional arguments) - use same code as everying else Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-02-03 17:20:52 +00:00
Markus Westerlind	0bd7941a18	fix(REPL): Don't buffer lines until a trailing semicolon is found and add history hinting (#3630 ) * fix(REPL): Don't buffer lines until a trailing semicolon is found The repl would silently buffer all lines until a trailing semicolon were found which resulted in some very confusing error messages as I would input invalid commands followed by a command I thought were valid, except I'd still get an error due to the previous command being buffered. This uses rustyline's helper feature to detect incomplete input (no trailing semicolon) and makes it accept multiline input until the input is completed. I also included some of rustyline's default hint and highlighting while I was at it. * chore: cargo clippy Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-02-03 17:11:01 +00:00
Marco Neumann	b3b2d9b623	feat: catalog setup CLI command (#3627 ) Closes #3509. Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-02-03 14:16:21 +00:00
Andrew Lamb	ab3c7573f5	test: add end to end for read_filter and empty string predicates (#3619 ) Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-02-03 14:05:32 +00:00
Marco Neumann	50cff27b01	chore: remove rdkafka dependency (#3625 ) All features are now covered by rskafka. This also removes the need to specify a server ID for write buffer consumers. This was only used for rdkafka since there we needed to specify a consumer group, even though we did not use any transactions.	2022-02-03 13:33:56 +00:00
kodiakhq[bot]	3197ea945b	Merge branch 'main' into dom/extract-ns-cache	2022-02-03 12:30:37 +00:00
Andrew Lamb	77b80e7618	fix(InfluxQL): treat null tags as `''` rather than `null` in storagerpc queries (#3557 ) * fix(InfluxQL): treat null tags as `''` rather than `null` in storage rpc queries * test: add one more case * fix: Update comment Co-authored-by: Raphael Taylor-Davies <1781103+tustvold@users.noreply.github.com> Co-authored-by: Raphael Taylor-Davies <1781103+tustvold@users.noreply.github.com> Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-02-03 12:14:43 +00:00
Paul Dix	ce46bbaada	feat: wire up the write buffer to the ingester process (#3533 ) This adds the scaffolding for the ingester server to consume data from Kafka. This ingests data in an in memory structure while creating records in the catalog for any partitions that don't yet exist. I've removed catalog_update.rs in ingester for now. That was mostly a placeholder and will be going in a combination of handler.rs and data.rs on my next PR which will have some primitive lifecycle wired up. There's one ugly bit here where the DML write is cloned because it's getting borrowed to output spans and metrics. I'll need to follow up with a refactor to make it so that the DML write's tables can be consumed without it gumming up the metrics stuff. Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-02-03 11:47:28 +00:00
Dom Dwyer	3cc4481616	refactor: extract NamespaceSchema cache Breaks the in-memory cache of NamespaceSchema out into a decoupled type that can be shared across multiple DML handlers.	2022-02-03 10:01:07 +00:00
Carol (Nichols \|\| Goulding)	a534136ccc	fix: Correct a 'long' argument name so the ingester command can run (#3621 ) Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-02-02 20:09:41 +00:00
Andrew Lamb	429d59f1b6	feat: Simplify predicates in the `InfluxRpcFrontend` before using them (#3588 ) * feat: normalize + simplify RPC predicates before using them * docs: Update predicate/src/rpc_predicate.rs Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-02-02 19:46:57 +00:00
kodiakhq[bot]	6de8ed4adc	Merge branch 'main' into dom/schema-validation	2022-02-02 16:05:41 +00:00
Luke Bond	6da15d9690	chore: cleanup catalog CLI output & args Co-authored-by: Marco Neumann <marco@crepererum.net>	2022-02-02 15:30:19 +00:00
Luke Bond	15827a534b	feat: catalog CLI command with update subcommand	2022-02-02 15:30:19 +00:00
Dom Dwyer	39d489d9e7	refactor: enable schema validation Adds the SchemaValidator to the DML handler stack - this adds it into the request path in router2.	2022-02-02 14:04:14 +00:00
Edd Robinson	5441682207	feat: add support for parsing predicate	2022-02-02 11:02:33 +00:00
Edd Robinson	08901c13cd	feat: support parsing timerange	2022-02-02 11:02:33 +00:00
Edd Robinson	a424d1c912	feat: shell command read_filter	2022-02-02 11:02:33 +00:00
Marco Neumann	59a2c74352	refactor: reusable ingester/router2 CLI pieces (#3590 ) * refactor: use a single CLI parser for ingester/router2 WB * refactor: reusable catalog DSN CLI handling We are going to need DSN handling for the router as well as for the some admin tools. * fix: DNS -> DSN	2022-02-01 12:57:58 +00:00
Marco Neumann	22778a3a80	chore: upgrade rskafka and parking_lot (#3592 )	2022-02-01 11:50:42 +00:00
Marco Neumann	b326b62b44	feat: buffer writes when writing to RSKafka (#3520 )	2022-02-01 10:07:52 +00:00
Carol (Nichols \|\| Goulding)	c633c9bc5c	feat: Wire object store into ingester persistence	2022-01-31 10:36:30 -05:00
Marco Neumann	c50fc8764d	feat: basic non-panic HTTP/gRPC interface for ingester (#3583 ) Don't panic when K8s requests a health status or someone requests a non-found HTTP route; or when we just TRY to start up the gRPC service.	2022-01-31 11:13:14 +00:00
Andrew Lamb	36642eb71d	test: add end to end tests that query missing tags (#3563 ) * test: add end to end tests that query missing tags * fix: add github reference Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-01-28 18:42:57 +00:00
Marco Neumann	3659a7f799	refactor: clean up ingester CLI (#3569 ) - use same args/envs names as router2 does - kafka => write buffer - add long forms to all CLI args so we don't have to pass positional arguments Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-01-28 17:58:16 +00:00
Raphael Taylor-Davies	4101d16f71	chore: feature flag consistency (#3574 ) * chore: feature flag consistency * chore: add aarch64-apple-darwin to hakari Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-01-28 16:38:59 +00:00
Marco Neumann	a22ca7c3d7	fix: router2 writer buffer topic (#3555 ) - Kafka does not support `_` in topic names, but `-` works, so let's change the default - Expose topic config via CLI/env	2022-01-28 10:10:04 +00:00
Andrew Lamb	f24ce03754	fix: provide correct environment variable to change log filter (#3561 ) Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-01-27 20:45:06 +00:00
Andrew Lamb	2062267d0f	chore: Update hashbrown (#3551 ) * chore: Update hashbrown * fix: hakari Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-01-27 15:34:10 +00:00
Dom	5447554aee	refactor(router2): DML handler stack (#3549 ) * refactor: composable DmlHandler stack Changes the DmlHandler trait to allow composition of handler logic in order to construct the complete request processing pipeline. * feat: debug log write/delete requests Log requests hitting the HTTP endpoint at DEBUG. * refactor: dml_handler -> dml_handlers Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-01-27 14:54:27 +00:00
Raphael Taylor-Davies	21c1824a7a	refactor: remove table_names from Predicate (#3545 ) * refactor: remove table_names from Predicate * chore: fix benchmarks * chore: review feedback Co-authored-by: Edd Robinson <me@edd.io> * chore: review feedback * chore: replace Default::default with InfluxRpcPredicate::default() Co-authored-by: Edd Robinson <me@edd.io> Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-01-27 14:44:49 +00:00
Paul Dix	16d584b2ff	feat: Add db_name/namespace to DmlWrite and DmlDelete (#3531 ) * feat: Add db_name/namespace to DmlWrite and DmlDelete This is required for the new ingester to be able to work with the write buffer. The protobuf that gets serialized over Kafka already includes the database name, it just wasn't getting carried through to the marshaled Dml operation. * fix: database != namespace, propagation through write buffer Co-authored-by: Marco Neumann <marco@crepererum.net> Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-01-27 14:12:20 +00:00
Andrew Lamb	5488c257d1	chore: Update datafusion, upgrade to arrow/parqet/arrow-flight 8.0.0 (#3517 ) * chore: Update datafusion * chore: update to arrow 8 * fix: update to use new DataFusion APIs * fix: update case for sortedness * fix: cargo hakari	2022-01-27 13:33:27 +00:00
Luke Bond	107f39d53c	feat: add trace collector to router2 (#3529 ) * feat: add trace collector to router2 * chore: fmt	2022-01-26 11:51:17 +00:00
Dom	6b0f7e6b2b	feat: initialise ShardedWriteBuffer (#3528 ) Initialises a ShardedWriteBuffer for the hard-coded "iox_shared" topic. Adds the following CLI flags: * --write-buffer: type of buffer [kafka, rskafka, file] * --write-buffer-addr: write buffer endpoint address The server uses these config options to initialise the appropriate write buffer backend, and configure the TableNamespaceSharder to shard operations over the set of sequencers exposed by the write buffer. Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-01-26 10:49:34 +00:00
Raphael Taylor-Davies	1b6aed063d	feat: add per-partition tracing (#3532 ) * feat: add per-partition tracing * chore: docs Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-01-26 10:39:21 +00:00
Raphael Taylor-Davies	db46ac04d0	feat: support line protocol precision parameter (#3522 ) (#3526 ) * feat: support line protocol precision parameter (#3522) * chore: format imports Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-01-25 14:12:22 +00:00
Paul Dix	bb893510a0	feat: Add scaffolding for ingester server * Adds a new ingester command to start an ingester server * Moves previous ingester server over to handler * Skeleton for gRPC and HTTP handlers	2022-01-21 18:02:19 -05:00
Andrew Lamb	9c19cd6cc4	fix: clamp start/end of TimestampRange to min/max valid timestamp values (#3487 ) Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-01-20 16:08:00 +00:00
Andrew Lamb	9751301374	refactor: rename `storage_api.rs` end to end test to `influxrpc.rs` for consistency (#3497 ) * refactor: rename `storage_api` end to end test to `influxrpc` for consistency * fix: fmt	2022-01-20 14:25:13 +00:00
Marco Neumann	168afb63ad	feat: add `size` methods to DML-related types This will be helpful when we want to batch DML operations in memory (e.g. when using RSKafka). This also ensures that `MBChunk` accounts for the column names that are stored within `MutableBatch`.	2022-01-18 13:52:31 +01:00
Dom	40a290f6f7	feat: router2 HTTP handlers Implements the HTTP v2 write API endpoint for router2.	2022-01-17 11:57:28 +00:00
Marco Neumann	c399e676ca	chore: upgrade clap to v3	2022-01-17 12:12:46 +01:00
Raphael Taylor-Davies	89db894df4	fix: serde_json `arbitrary_precision` (#3458 ) (#3469 ) Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-01-17 09:50:54 +00:00
Andrew Lamb	b036db293f	fix: Format `ParenExpression` RPC storage Expression `Node`s (#3463 ) Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-01-14 16:02:51 +00:00
Edd Robinson	36ec6019f9	feat: wire up token to query frontends	2022-01-14 10:26:11 +00:00
Andrew Lamb	dd23056efd	chore: update datafusion, arrow, prost, tonic, pbjson, etc (#3455 ) * chore: update datafusion, arrow, prost, tonic, etc * fix: update pprof as well * chore: update hakari * fix: update pbjson * chore: update heappy * fix: hakari * fix: workaround https://github.com/influxdata/influxdb_iox/issues/3458 Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-01-13 17:07:15 +00:00
Dom	430823c148	docs: fix typo Co-authored-by: Marko Mikulicic <mkm@influxdata.com>	2022-01-13 15:42:30 +00:00
Dom	3d32901877	fix: undected gzip HTTP body truncation When reading the gzip-encoded body of a HTTP request, the stream is read up until the configured maximum number of allowable bytes, at which point the body was silently trucated. This could allow fields in submitted line protocol to be silently lost (amongst other bad things). This change ensures that truncation results in a RequestSizeExceeded error.	2022-01-13 14:38:37 +00:00
Dom	7fc17203e2	refactor: add router2 server mode Plumbs the router2 crate into IOx's CLI & server-runner framework.	2022-01-12 14:52:47 +00:00
Dom	a8cb8755de	feat: new router2 crate This commit adds an almost-empty router2 crate containing enough of a skeleton to plumb into the IOx CLI/server runner.	2022-01-12 14:43:10 +00:00
Marco Neumann	f3f6f335a9	chore: upgrade to snafu 0.7 (#3440 )	2022-01-11 19:22:36 +00:00
Andrew Lamb	b76921d26e	fix(influxrpc): Support _field references in front end conversion (#3426 ) Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-01-10 21:54:56 +00:00
Andrew Lamb	6d20ce1f9e	feat: Allow wipe catalog in `ReplayError` and `WriteBufferCreationError` states (#3425 ) * feat: feat: Allow wipe catalog in ReplayError * fix: comments Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-01-07 17:07:44 +00:00
Andrew Lamb	336ffd1966	refactor: Remove `Result` in QueryDatabase trait (none of the functions can fail) (#3422 ) Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2022-01-06 22:03:08 +00:00
Andrew Lamb	a93ae739a9	feat: Add table_name to Partition API (#3421 )	2022-01-06 16:38:39 +00:00
Carol (Nichols \|\| Goulding)	f9174c483b	refactor: Extract server::db into its own crate (#3417 ) * refactor: Extract JobRegistry from the server crate Both the server crate and a db crate that I'm about to extract depend on JobRegistry, so to avoid making circular dependencies, extract the JobRegistry to its own crate. * refactor: Move db out of server into its own crate Fixes #2821.	2021-12-23 22:01:17 +00:00
Carol (Nichols \|\| Goulding)	2c3ca0c77c	docs: Add doc comments to all CLI subcommands (#3414 ) * docs: Add doc comments to all CLI subcommands * docs: Update influxdb_iox/src/commands/database/recover.rs Co-authored-by: Andrew Lamb <alamb@influxdata.com> Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2021-12-23 12:17:39 +00:00
Carol (Nichols \|\| Goulding)	5aaee1bcf4	fix: Remove some straggling references to writer ID that should be server ID (#3415 ) Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2021-12-22 21:05:38 +00:00
Edd Robinson	c631dd4267	refactor: address PR feedback	2021-12-21 10:23:12 +00:00
Edd Robinson	200c32cb32	feat: support all tag key predicates	2021-12-21 10:23:12 +00:00
Edd Robinson	5b277c8754	test: exercise show tag values	2021-12-17 17:59:30 +00:00
Edd Robinson	119e579fa3	feat: add support for tag values grouped by measurement and tag key	2021-12-17 17:59:30 +00:00
Edd Robinson	a68d71ef06	feat: add ability to materialise measurements	2021-12-17 17:59:30 +00:00
Edd Robinson	ebbf86dc1d	refactor: impl stub	2021-12-17 17:59:30 +00:00
Edd Robinson	3fde298a99	refactor: handle rpc	2021-12-17 17:59:30 +00:00
Marco Neumann	9f2694bf1b	test: `objest_store` azure support via Azurite	2021-12-17 09:30:21 +01:00
Andrew Lamb	64f915d860	fix: flaky end to end system_tables test (#3397 )	2021-12-17 08:13:09 +00:00
Marco Neumann	5d58b06e64	test: fix some environment variables influencing our tests	2021-12-16 11:17:36 +01:00
Marco Neumann	d8810074e8	refactor: better test action CLI	2021-12-16 09:29:25 +01:00
Marco Neumann	f4fde15810	refactor: nicer code Co-authored-by: Carol (Nichols \|\| Goulding) <193874+carols10cents@users.noreply.github.com>	2021-12-16 09:29:25 +01:00
Marco Neumann	ed775431b6	fix: treat early server worker exit as proper error Instead of just logging the issue, also make sure the error gets propagated all the way up to the exit code. Fixes #3375.	2021-12-16 09:29:25 +01:00
Andrew Lamb	da0330fc5f	fix: print newline at end of `database list --detailed` (#3382 )	2021-12-15 19:33:24 +00:00
Andrew Lamb	758b65dd29	feat: Add database initialization state and errors to CLI and remove list_databases_detailed gRPC (#3377 ) * feat: Add database initialization state and errors to CLI: * fix: do not use optional in protobuf * fix: clippy * fix: correct check I broke appeasing clippy	2021-12-15 12:18:41 +00:00
Edd Robinson	7fe6897c59	refactor: add support for handling `TagValuesGroupedByMeasurementAndTagKeyRequest` (#3373 ) * chore: update Storage service * refactor: handle rpc * refactor: log request fields * refactor: add new RPC to client * test: basic test of RPC * refactor: fmt Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2021-12-14 18:42:12 +00:00
Edd Robinson	ec6d945cbe	feat: wire up general predicates to measurement_names	2021-12-13 14:01:09 +00:00
Nga Tran	c0ba69f09e	chore: marge main to branch and resolve conflict	2021-12-09 15:40:33 -05:00
Nga Tran	35370922f3	refactor: make a setup for 2 persisted chunks that can be used in for different places	2021-12-09 15:21:56 -05:00
Andrew Lamb	3cda6b6c0f	refactor: Remove collect_query and replication (#3348 ) Co-authored-by: kodiakhq[bot] <49736102+kodiakhq[bot]@users.noreply.github.com>	2021-12-09 19:58:19 +00:00
Nga Tran	099e2d4056	chore: Apply suggestions from code review Co-authored-by: Marco Neumann <marco@crepererum.net>	2021-12-09 13:52:42 -05:00
Nga Tran	e46708354e	test: add management cli tests	2021-12-09 12:53:45 -05:00

1 2 3 4 5 ...

457 Commits (5d66cd0a81d2533ee174e1bfb56bab86f81c1fed)