influxdb

Commit Graph

Author	SHA1	Message	Date
Trevor Hilton	b7fd8e2386	feat: remove metadata caches on db and table delete (#25599 )	2024-11-28 11:35:29 -05:00
Trevor Hilton	81715fbfea	refactor: display column names for predicates in EXPLAIN for metadata cache (#25598 )	2024-11-28 11:18:12 -05:00
Trevor Hilton	13ab41fa1f	feat: CLI to create and delete metadata caches (#25595 ) This adds two new CLI commands to the `influxdb3` binary: * `influxdb3 meta-cache create` * `influxdb3 meta-cache delete` To create and delete metadata caches, respectively. A basic integration test was added to check that this works E2E. The `influxdb3_client` was updated with methods to create and delete metadata caches, and which is what the CLI commands use under the hood.	2024-11-28 09:04:20 -05:00
Trevor Hilton	9ead1dfe4b	feat: meta_caches system table (#25593 ) This adds a new system table "meta_caches" that allows users to view the state of their metadata caches on a per-db basis An integration test was added to verify that it works.	2024-11-28 08:57:02 -05:00
Trevor Hilton	234d37329a	feat: metacache REST APIs to create and delete (#25587 )	2024-11-27 08:41:46 -05:00
praveen-influx	bfa0e71558	feat: make query executor as trait object (#25591 ) * feat: make query executor as trait object This commit moves `QueryExecutorImpl` behind a `dyn` (trait object) as we have other impls in core for `QueryExecutor` and this will keep both pro and OSS traits in sync * chore: fix cargo audit failures - address https://rustsec.org/advisories/RUSTSEC-2024-0399.html by running `cargo update --precise 0.23.18 --package rustls@0.23.14` - address yanked version of `url` crate (2.5.3) by running `cargo update -p url`	2024-11-26 17:18:22 +00:00
Trevor Hilton	8e23032ceb	feat: add metadata cache provider with APIs for write and query (#25566 ) This adds the MetaDataCacheProvider for managing metadata caches in the influxdb3 instance. This includes APIs to create caches through the WAL as well as from a catalog on initialization, to write data into the managed caches, and to query data out of them. The query side is fairly involved, relying on Datafusion's TableFunctionImpl and TableProvider traits to make querying the cache using a user-defined table function (UDTF) possible. The predicate code was modified to only support two kinds of predicates: IN and NOT IN, which simplifies the code, and maps nicely with the DataFusion LiteralGuarantee which we leverage to derive the predicates from the incoming queries. A custom ExecutionPlan implementation was added specifically for the metadata cache that can report the predicates that are pushed down to the cache during query planning/execution. A big set of tests was added to to check that queries are working, and that predicates are being pushed down properly.	2024-11-22 10:57:26 -05:00
praveen-influx	3cde24feb4	feat: delete table (#25572 ) This commit allows deleting (soft) a table. For an user, following command will allow soft deleting a table (bar) in db (foo) ``` influxdb3 table delete --dbname foo --table bar --host $host ``` - Added `soft_delete_table` to `DatabaseManager` trait, which already hosts `soft_delete_database` method. The code roughly follows the same flow as db delete. Although like db schema, it does clone on write because the reference is behind an Arc, `Arc::make_mut` is used in this change. - Moved db delete related cli parser under "manage" module that has both db and table delete functionality - Some minor tidyups (removing unused methods, renaming method so that the order in name matches actual return type eg. `table_id_and_schema`, should return (id, schema) and not (schema, id)) closes: https://github.com/influxdata/influxdb/issues/25561	2024-11-22 08:42:45 +00:00
Jackson Newhouse	956e223388	fix: don't rebuild snapshot if it has already been taken. (#25570 )	2024-11-20 08:55:42 -08:00
Michael Gattozzi	230bf02f93	feat: delete old Catalogs on persist (#25568 ) This commit changes the code so that we only keep the 10 most recent Catalogs. When a new one is persisted we delete any old ones that exist. If the deletion would fail we don't panic and let a future persist cleanup the catalogs rather than failing the persist itself. This commit also adds a test to make sure that only the catalogs we expect to are deleted on persist.	2024-11-19 12:42:20 -05:00
praveen-influx	33c2d47ba9	feat: drop/delete database (#25549 ) * feat: drop/delete database This commit allows soft deletion of database using `influxdb3 database delete <db_name>` command. The write buffer and last value cache are cleared as well. closes: https://github.com/influxdata/influxdb/issues/25523 * feat: reuse same code path when deleting database - In previous commit, the deletion of database immediately triggered clearing last cache and query buffer. But on restarts same logic had to be repeated to allow deleting database when starting up. This commit removes immediate deletion by explicitly calling necessary methods and moves the logic to `apply_catalog_batch` which already applies `CatalogOp` and also clearing cache and buffer in `buffer_ops` method which has hooks to call other places. closes: https://github.com/influxdata/influxdb/issues/25523 * feat: use reqwest query api for query param Co-authored-by: Trevor Hilton <thilton@influxdata.com> * feat: include deleted flag in DatabaseSnapshot - `DatabaseSchema` serialization/deserialization is delegated to `DatabaseSnapshot`, so the `deleted` flag should be included in `DatabaseSnapshot` as well. - insta test snapshots fixed closes: https://github.com/influxdata/influxdb/issues/25523 * feat: address PR comments + tidy ups --------- Co-authored-by: Trevor Hilton <thilton@influxdata.com>	2024-11-19 16:08:14 +00:00
Trevor Hilton	53f54a6845	feat: metadata cache core impl (#25552 ) * feat: core metadata cache structs with basic tests Implement the base MetaCache type that holds the hierarchical structure of the metadata cache providing methods to create and push rows from the WAL into the cache. Added a prune method as well as a method for gathering record batches from a meta cache. A test was added to check the latter for various predicates and that the former works, though, pruning shows that we need to modify how record batches are produced such that expired entries are not emitted. * refactor: filter expired entries and do some clean up in the meta cache	2024-11-18 12:28:12 -05:00
Trevor Hilton	2ac3df1bca	refactor: use `SerdeVecMap` in `PersistedSnapshot` (#25541 ) * refactor: use SerdeVecMap in PersistedSnapshot This changes from the use of a HashMap to store the DB -> Table structure in the PersistedSnapshot files to using a SerdeVecMap, which will have the identifiers serialized as integers instead of strings. * test: add a snapshot test for persisted snapshots	2024-11-12 16:31:36 -05:00
praveen-influx	814eb31309	chore: update core deps (#25532 ) * chore: update core deps - arrow/parquet deps are patched (as in core) - three specific code changes to cope with changes in core crates - TransitionPartitionId, use `from_parts` instead of `new` - arrow buffers can take &[u8] directly without `to_vec()`/`vec!` (used only in tests) - `schema` and `influxdb_line_protocol` crates need `v3` feature enabled * chore: update deny.toml * chore: formatting and deny toml changes Unicode-3.0 license is added to allowed licenses list, without it end up with 19 errors (`zerovec`, `zerovec-derive` etc.) * chore: address PR feedback - move enabling v3 feature to root Cargo.toml - added the upstream PR for datafusion-common that introduced RUSTSEC-2024-0384	2024-11-12 16:07:31 +00:00
Paul Dix	35e29d1408	feat: Update catalog to use sequence number in path (#25526 ) Updates the catalog to use its own sequence number in the path. This will enable downstream Pro systems that pick up PersistedSnapshots to get the specific catalog that a snapshot is associated with since its sequence number is included. Also updated the type to be CatalogSequenceNumber to make it more clear & readable when being used.	2024-11-08 15:50:15 -05:00
Trevor Hilton	3bb63b2d71	fix: throw error when adding fields to non-existent table in WAL (#25525 ) * fix: throw error when adding fields to non-existent table * test: add test for expected behaviour in catalog op apply This also added in some helpers to the wal crate that were previously added to pro.	2024-11-08 13:15:07 -05:00
praveen-influx	c2b8a3a355	feat: add column names to last cache sys table (#25521 ) * feat: add column names to last cache sys table closes: https://github.com/influxdata/influxdb/issues/25511 * feat: move all `get_by_id` methods to take reference in schema	2024-11-08 16:08:30 +00:00
Trevor Hilton	391b67f9ab	feat: configurable last cache eviction (#25520 ) * refactor: make last cache eviction optional This changes how the last cache is evicted. It will no longer run eviction on writes to the cache, instead, there is an optional method to create a last cache provider that will run eviction in a background task on a specified interval. Otherwise, when records are produced from the cache, only those that have not expired will be produced. This should reduce locks on the cache and hopefully improve performance. * feat: configurable last cache eviction interval * docs: clean up var names, code docs, and comments	2024-11-06 09:59:17 -05:00
Trevor Hilton	da294a265e	refactor: remove conversion step in validator (#25515 ) The write validator now builds the row data destined for the WAL directly, vs. creating an intermediate type to hold row data.	2024-11-04 15:23:54 -05:00
Trevor Hilton	ec01934c57	chore: remove unnecessary rustsec for the tonic cve (#25516 ) `cargo deny` was showing that no crate matched the advisory criteria for this [RUSTSEC advisory](https://rustsec.org/advisories/RUSTSEC-2024-0376.html), so this PR removes the ignore entry. In addition, the `hashbrown` crate was causing a new audit failure, and updating it required that the `Zlib` license be added to our list of allowed licenses. No issue for this, but it is blocking another PR at the moment (https://github.com/influxdata/influxdb/pull/25515).	2024-11-04 15:02:39 -05:00
praveen-influx	f745ea69c8	fix: default url for telemetry should include protocol (#25514 ) closes: https://github.com/influxdata/influxdb/issues/25502	2024-11-04 14:20:54 +00:00
Trevor Hilton	aa70a73487	fix: the cache target for build artefacts in Dockerfile (#25510 )	2024-11-01 17:20:00 -04:00
Trevor Hilton	5698e79a34	feat: helper methods on WalOp (#25486 )	2024-11-01 17:19:20 -04:00
Trevor Hilton	d26a73802a	refactor: move to `ColumnId` and `Arc<str>` as much as possible (#25495 ) Closes #25461 _Note: the first three commits on this PR are from https://github.com/influxdata/influxdb/pull/25492_ This PR makes the switch from using names for columns to the use of `ColumnId`s. Where column names are used, they are represented as `Arc<str>`. This impacts most components of the system, and the result is a fairly sizeable change set. The area where the most refactoring was needed was in the last-n-value cache. One of the themes of this PR is to rely less on the arrow `Schema` for handling the column-level information, and tracking that info in our own `ColumnDefinition` type, which captures the `ColumnId`. I will summarize the various changes in the PR below, and also leave some comments in-line in the PR. ## Switch to `u32` for `ColumnId` The `ColumnId` now follows the `DbId` and `TableId`, and uses a globally unique `u32` to identify all columns in the database. This was a change from using a `u16` that was only unique within the column's table. This makes it easier to follow the patterns used for creating the other identifier types when dealing with columns, and should reduce the burden of having to manage the state of a table-scoped identifier. ## Changes in the WAL/Catalog * `WriteBatch` now contains no names for tables or columns and purely uses IDs * This PR relies on `IndexMap` for `_Id`-keyed maps so that the order of elements in the map is consistent. This has important implications, namely, that when iterating over an ID map, the elements therein will always be produced in the same order which allows us to make assertions on column order in a lot of our tests, and allows for the re-introduction of `insta` snapshots for serialization tests. This map type provides O(1) lookups, but also provides _fast_ iteration, which should help when serializing these maps in write batches to the WAL. * Removed the need to serialize the bi-directional maps for `DatabaseSchema`/`TableDefinition` via use of `SerdeVecMap` (see comments in-line) * The `tables` map in `DatabaseSchema` no stores an `Arc<TableDefinition>` so that the table definition can be shared around more easily. This meant that changes to tables in the catalog need to do a clone, but we were already having to do a clone for changes to the DB schema. * Removal of the `TableSchema` type and consolidation of its parts/functions directly onto `TableDefinition` * Added the `ColumnDefinition` type, which represents all we need to know about a column, and is used in place of the Arrow `Schema` for column-level meta-info. We were previously relying heavily on the `Schema` for iterating over columns, accessing data types, etc., but this gives us an API that we have more control over for our needs. The `Schema` is still held at the `TableDefinition` level, as it is needed for the query path, and is maintained to be consistent with what is contained in the `ColumnDefinition`s for a table. ## Changes in the Last-N-Value Cache * There is a bigger distinction between caches that have an explicit set of value columns, and those that accept new fields. The former should be more performant. * The Arrow `Schema` is managed differently now: it used to be updated more than it needed to be, and now is only updated when a row with new fields is pushed to a cache that accepts new fields. ## Changes in the write-path * When ingesting, during validation, field names are qualified to their associated column ID	2024-11-01 16:42:57 -04:00
Trevor Hilton	0e814f5d52	feat: SerdeVecMap type for serializing ID maps (#25492 ) This PR introduces a new type `SerdeVecHashMap` that can be used in places where we need a HashMap with the following properties: 1. When serialized, it is serialized as a list of key-value pairs, instead of a map 2. When deserialized, it assumes the serialization format from (1.) and deserializes from a list of key-value pairs to a map 3. Does not allow for duplicate keys on deserialization This is useful in places where we need to create map types that map from an identifier (integer) to some value, and need to serialize that data. For example: in the WAL when serializing write batches, and in the catalog when serializing the database/table schema. This PR refactors the code in `influxdb3_wal` and `influxdb3_catalog` to use the new type for maps that use `DbId` and `TableId` as the key. Follow on work can give the same treatment to `ColumnId` based maps once that is fully worked out. ## Explanation If we have a `HashMap<u32, String>`, `serde_json` will serialize it in the following way: ```json {"0": "foo", "1": "bar"} ``` i.e., the integer keys are serialized as strings, since JSON doesn't support any other type of key in maps. `SerdeVecHashMap<u32, String>` will be serialized by `serde_json` in the following way: ```json, [[0, "foo"], [1, "bar"]] ``` and will deserialize from that vector-based structure back to the map. This allows serialization/deserialization to run directly off of the `HashMap`'s `Iterator`/`FromIterator` implementations. ## The Controversial Part One thing I also did in this PR was switch the catalog from using a `BTreeMap` for tables to using the new `HashMap` type. This breaks the deterministic ordering of the database schema's `tables` map and therefore wrecks the snapshot tests we were using. I had to comment those parts of their respective tests out, because there isn't an easy way to make the underlying hashmap have a deterministic ordering just in tests that I am aware of. If we think that using `BTreeMap` in the catalog is okay over a `HashMap`, then I think it would be okay to roll a similar `SerdeVecBTreeMap` type specifically for the catalog. Coincidentally, this may actually be a good use case for [`indexmap`](https://docs.rs/indexmap/latest/indexmap/), since it holds supposedly similar lookup performance characteristics to hashmap, while preserving order and _having faster iteration_ which could be a win for WAL serialization speed. It also accepts different hashing algorithms so could be swapped in with FNV like `HashMap` can. ## Follow-up work Use the `SerdeVecHashMap` for column data in the WAL following https://github.com/influxdata/influxdb/issues/25461	2024-10-25 13:49:02 -04:00
praveen-influx	cdb1e862b9	feat: set tcp_nodelay in hyper builder (#25488 ) closes: https://github.com/influxdata/influxdb/issues/25466	2024-10-24 21:48:24 +01:00
Trevor Hilton	f1bd868d78	fix: last cache cli parser bug (#25485 ) Correctly parse key and value column args to create last cache command	2024-10-23 12:30:05 -04:00
Trevor Hilton	2eb2341599	refactor: method on catalog did not need mut self (#25483 )	2024-10-22 14:57:57 -04:00
Trevor Hilton	afd0022210	feat: add insta crate yaml feature (#25482 )	2024-10-22 12:30:28 -04:00
praveen-influx	4e6c5dc825	feat: expose path used by cache request (#25480 )	2024-10-22 14:44:44 +01:00
Trevor Hilton	ce9276d96d	refactor: changes needed for IDs in pro (#25479 ) * refactor: roll back addition of DatabaseSchemaProvider trait * refactor: make parquet metrics optional in telemetry for pro * refactor: make ParquetFileId Hash * refactor: test harness logging	2024-10-21 15:17:02 -04:00
praveen-influx	5473a9489b	feat: allow telemetry endpoint to be passed in (#25475 ) Allow the endpoint for telemetry to be passed in via the cli args, e.g ``` --telemetry-endpoint "https://somehost/test/" ``` and the actual endpoint always appends `v3` to it. So, above URL becomes "https://somehost/test/v3"	2024-10-18 14:34:27 +01:00
Trevor Hilton	10b6a2810d	refactor: separte catalog schema API (#25468 ) Separate out methods of the Catalog API that are used on the query side into a new trait `DatabaseSchemaProvider`. The new trait includes methods from the Catalog that get the underlying `DatabaseSchema` or interact with names/IDs. This will allow for a separate implementation of the Catalog for pro that only needs to hold a replicated/combined view in-memory of one or more catalogs without the need to do persistence that a write buffer's catalog needs to do. While in there I also switched the `QueryExecutorImpl::new` method to take an args struct to avoid the clippy lint.	2024-10-16 11:42:07 -04:00
Trevor Hilton	ead6d27c5f	refactor: improvements to the catlog API (#25463 )	2024-10-11 13:44:43 -04:00
Michael Gattozzi	eb24b3bc07	fix: lint fixes for #25388 (#25451 )	2024-10-10 17:11:27 -04:00
Michael Gattozzi	724a7e99c3	feat: Add non-unique u16 Id to ColumnDefinition (#25388 ) * feat: Add non-unique u16 Id to ColumnDefinition This commit adds the column_id field to ColumnDefinition so that the output for a Catalog will contain the id of that column. This is non unique, whereas TableIds and DbIds will be unique. The column_id corresponds to it's index in the schema. Closes #25386	2024-10-10 16:59:10 -04:00
Michael Gattozzi	219fc168ea	refactor: Move ParquetFileId into influxdb3_id (#25449 )	2024-10-10 12:07:15 -04:00
Jamie Strandboge	277a6e0a87	chore(circleci): update to use language-checker (#25443 )	2024-10-10 08:17:04 -05:00
praveen-influx	87f198a68b	fix: check num items to prune before pruning parquet cache (#25447 ) When running the tests repeatedly the tests failed intermittently as the background runner wakes up to prune the cache and the tests are loading and removing the data. The check whether `n_to_prune` is greater than 0 before going ahead with pruning fixes the issue Closes: https://github.com/influxdata/influxdb/issues/25446	2024-10-10 14:03:26 +01:00
Jamie Strandboge	0835093c78	feat(circleci): add inclusivity checks (#25437 ) * feat(circleci): add inclusivity checks * chore(circleci): adjust package-validation for inclusive language * chore: update tests for inclusive language	2024-10-09 08:01:31 -05:00
praveen-influx	1f1125c767	refactor: udpate docs and tests for the telemetry crate (#25432 ) - Introduced traits, `ParquetMetrics` and `SystemInfoProvider` to enable writing easier tests - Uses mockito for code that depends on reqwest::Client and also uses mockall to generally mock any traits like `SystemInfoProvider` - Minor updates to docs	2024-10-08 15:45:13 +01:00
Michael Gattozzi	c4534b06da	feat: move Table Id/Name mapping into DB Schema (#25436 )	2024-10-08 10:03:55 -04:00
praveen-influx	bd20f80ce2	chore: udpate docs for the telemetry crate (#25431 ) * chore: udpate docs for the telemetry crate * chore: Update influxdb3_telemetry/src/stats.rs Co-authored-by: Michael Gattozzi <mgattozzi@influxdata.com> * chore: Update influxdb3_telemetry/src/sender.rs Co-authored-by: Michael Gattozzi <mgattozzi@influxdata.com> --------- Co-authored-by: Michael Gattozzi <mgattozzi@influxdata.com>	2024-10-08 09:24:49 +01:00
Trevor Hilton	42672e06b0	chore: cargo update (#25433 )	2024-10-07 10:50:53 -04:00
Trevor Hilton	533ebff1d7	fix: add host id to parquet file paths (#25428 )	2024-10-04 07:44:42 -04:00
Michael Gattozzi	eeb1aa7905	feat: swap over to DbId and TableId everywhere (#25421 ) * feat: Add TableId and ColumnId * feat: swap over to DbId and TableId everywhere This commit swaps us over to using the DbId and TableId types everywhere for our internal systems. Anywhere that's external facing, such as names for last cache tables or line protocol parsing, use names. In these cases we have the `Catalog` which keeps a map of TableIds and DbIds in a bidirectional mapping for easy lookup i.e. id <-> names. While in essence the change itself isn't that complicated given the nature of how much we depended on names for things, the changes end up being quite invasive and extensive. Luckily it shouldn't be too hard to review. Note this does not add the column ids which will be done in a follow up PR. Closes #25375 Closes #25403 Closes #25404 Closes #25405 Closes #25412 Closes #25413	2024-10-03 14:47:46 -04:00
Trevor Hilton	1cd930fec1	fix: flaky parquet cache for real this time (#25426 ) This adds a watch channel to check for when prunes have happened on the parquet cache oracle, so we can notify something, like a test, that needs to know when a prune has happened. This should make the cache eviction test in the parquet_cache module not flake out anymore.	2024-10-03 10:59:27 -04:00
praveen-influx	8ccb580162	feat: telemetry report for parquet metrics (#25425 ) - added mechanism within PersistedFile to expose parquet file related metrics. The details are updated when new snapshot is generated and also when all snapshots are loaded when the process starts up - at the point of creating the telemetry payload these parquet metrics are looked up before sending it to the server. Closes: https://github.com/influxdata/influxdb/issues/25418	2024-10-03 15:11:40 +01:00
Trevor Hilton	7d37bbbce7	test: add test helpers for object store types (#25420 ) This adds a new crate `influxdb3_test_helpers` which provides two object store helper types that can be used to track request counts made through the store, as well as synchronize requests made through the store, resp.	2024-10-02 14:45:12 -04:00
Trevor Hilton	7a903ca080	chore: unblock CI for tonic/hyper audit failures (#25419 ) We will need to wait on the RUSTSEC advisory to be resolved upstream, i.e., by having tonic and hyper upgraded in core, before we can lift this advisory ignore and use the latest versions of those crates.	2024-10-02 14:09:20 -04:00

1 2 3 4 5 ...

49432 Commits (7ea72752a642ced80cb39e30cff42a9a90a189d0) All Branches Search

49432 Commits (7ea72752a642ced80cb39e30cff42a9a90a189d0)

All Branches