mentat

Author	SHA1	Message	Date
Emily Toop	e1e7cbaa44	Closes #634 - Fix variables in predicates (#635 ) r=rnewman We were forgetting to check for bound variables when resolving types other than ref types during inequality handling. This patch adds in the binding checks and `bails` if the bound variable is of the wrong type. #634	2018-05-09 16:24:12 +01:00
Richard Newman	e21156a754	Implement simple pull expressions (#638 ) r=nalexander * Refactor AttributeCache populator code for use from pull. * Pre: add to_value_rc to Cloned. * Pre: add From<StructuredMap> for Binding. * Pre: clarify Store::open_empty. * Pre: StructuredMap cleanup. * Pre: clean up a doc test. * Split projector crate. Pass schema to projector. * CLI support for printing bindings. * Add and use ConjoiningClauses::derive_types_from_find_spec. * Define pull types. * Implement pull on top of the attribute cache layer. * Add pull support to the projector. * Parse pull expressions. * Add simple pull support to connection objects. * Tests for pull. * Compile with Rust 1.25. The only choice involved in this commit is that of replacing the anonymous lifetime '_ with a named lifetime for the cache; since we're accepting a Known, which includes the cache in question, I think it's clear that we expect the function to apply to any given cache lifetime. * Review comments. * Bail on unnamed attribute. * Make assert_parse_failure_contains safe to use. * Rework query parser to report better errors for pull. * Test for mixed wildcard and simple attribute.	2018-05-04 12:56:00 -07:00
Nick Alexander	90465ae74a	Flip ValueRc to Arc in order to allow TypedValue in errors. (#677 ) (#678 ) r=rnewman @mmacedoeu did a good deal of work to show that Arc instead of Rc wasn't too difficult in #656, and @rnewman pushed the refactoring across the line in #659. However, we didn't flip the switch at that time. For #673, we'd like to include TypedValue instances in errors, and with error-chain (and failure) error types need to be 'Sync + 'Send, so we need Arc. This builds on #659 and should also finish #656.	2018-05-03 16:46:49 -07:00
Richard Newman	f979044ba1	Refactor value type boxing. (#659 ) r=nalexander * Pre: eliminate some occurrences of Rc, largely through the magic of Into. * Pre: introduce FromRc to convert between refcounted types. * Introduce ValueRc as an abstraction over Rc/Arc choice. * Move Cloned to core. * Move CString-creation methods to TypedValue. * Finish transition.	2018-04-25 14:23:27 -07:00
Richard Newman	a2e13f624c	Add 'Binding', a structured value type to return from queries. (#657 ) r=nalexander Bump to 0.7: breaking change.	2018-04-24 15:08:38 -07:00
Richard Newman	1818e0b98e	Split mentat_core TypedValue code into separate files for clarity.	2018-04-24 15:05:04 -07:00
Richard Newman	a74a2deffc	Introduce RelResult rather than Vec<Vec<TypedValue>>. (#639 ) r=nalexander * Pre: clean up core/src/lib.rs. * Pre: use indexmap 1.0 in db and query-projector. * Change rel results to be a RelResult instance, not a Vec<Vec<TypedValue>>. This avoids memory fragmentation and improves locality by using a single heap-allocated vector for all bindings, rather than a separate heap-allocated vector for each row. We hide this abstraction behind the `RelResult` type, which tracks the stride length (width) of each row. * Don't allocate temporary vectors when projecting RelResults.	2018-04-24 15:04:00 -07:00
Richard Newman	3f8464e8ed	Implement vocabulary-driven schema upgrades. (#595 ) r=emily	2018-04-09 09:47:49 -07:00
Emily Toop	b3e27d86a9	Add converter functions from TypedValue to underlying type	2018-04-06 10:46:15 +01:00
Richard Newman	8607ecb745	Fix merge conflict.	2018-04-03 14:54:46 -07:00
Richard Newman	4d8e179a59	Expose component_attributes on Schema. (#623 ) r=nalexander Some parts of the query engine and transactor need to know whether an attribute is a component attribute, and sometimes want to do so in a generated SQL query. This is one way to do that.	2018-04-03 14:25:53 -07:00
Richard Newman	558906df4f	Fix: db/component should be db/isComponent. (#624 ) r=nalexander	2018-04-03 14:25:28 -07:00
Richard Newman	6c54e1d370	Support :db/noHistory for attributes. (#622 ) r=nalexander At this point we never discard history, but this completes the API support for doing so.	2018-04-03 14:23:46 -07:00
Richard Newman	833ff92436	Simple aggregates. (#584 ) r=emily * Pre: use debugcli in VSCode. * Pre: wrap subqueries in parentheses in output SQL. * Pre: add ExistingColumn. This lets us make reference to columns by name, rather than only pointing to qualified aliases. * Pre: add Into for &str to TypedValue. * Pre: add Store.transact. * Pre: cleanup. * Parse and algebrize simple aggregates. (#312) * Follow-up: print aggregate columns more neatly in the CLI. * Useful ValueTypeSet helpers. * Allow for entity inequalities. * Add 'differ', which is a ref-specialized not-equals. * Add 'unpermute', a function for getting unique, distinct pairs from bindings. * Review comments. * Add 'the' pseudo-aggregation operator. This allows for a corresponding value to be returned when a query includes one 'min' or 'max' aggregate.	2018-03-12 15:18:50 -07:00
Richard Newman	f42ae35b70	Update cache on write. (#566 ) r=emily * Use the cache to make constant queries super fast. * Fix translate tests to match: we no longer generate SQL for many of them! * Accumulate additions and removals into the cache. * Make attribute cache clone-on-write; store it in Metadata. * Allow caching of fulltext attributes, interning strings.	2018-03-06 09:01:20 -08:00
Richard Newman	e33fe71c47	Rework caching and use it inside the query engine. (#553 ) r=emily This puts caching in mentat_db, adds a reverse lookup capability for unique attributes, and populates bidirectional caches with a single SQL cursor walk. Differentiate between begin_read and begin_uncached_read. Note that we still allow toggling within InProgress, because there might be transient local state that makes starting a new transaction impossible.	2018-02-21 11:51:45 -08:00
Grisha Kruglov	84f29676e8	"Unchanged server" uploader flow (#543 ) r=rnewman * Remove unused struct from tx_processor * Derive serialize & deserialize for TypedValue * First pass of uploader flow + feedback	2018-02-09 09:55:19 -08:00
Richard Newman	2614f498be	Ergonomics improvements, including a `kw` macro. (#537 ) r=emily * Add TypedValue::instant(micros). * Add From<f64> for TypedValue. * Add lookup_values_for_attribute to Conn. * Add q_explain to Queryable. * Expose an iterator over FindSpec's columns. * Export edn from mentat crate. Export QueryExecutionResult. * Implement Display for Variable and Element. * Introduce a `kw` macro. This allows you to write: ```rust kw!(:foo/bar) ``` instead of ```rust NamespacedKeyword::new("foo", "bar") ``` … and it's more efficient, too. Add `mentat::open`, eliminate use of `mentat_db` in some places.	2018-02-01 09:27:23 -08:00
Edouard Oger	3bf7459315	Allow customers to assert facts about the current transaction. (#225 ) r=rnewman Also move `now` into core, implement microsecond truncation. This is so we don't return a more granular -- and thus subtly different -- timestamp in a `TxReport` than we put into the store.	2018-01-29 16:46:04 -08:00
Thom Chiovoloni	98502eb68f	Implement type annotations in queries. (#526 ) r=rnewman	2018-01-29 14:37:53 -08:00
Richard Newman	4acc6d0658	InProgressRead, KnownEntid. r=nalexander,emily Improve naming of read-only transactions. Implement entid_for_type. Simplify get_attribute. Name ignored var in algebrizer. Comment attribute_for_ident. Make KnownEntid a core concept. Expose lookup_value_for_attribute. Implement HasSchema and a new query encapsulation on Conn. Pre: export Queryable.	2018-01-23 08:40:18 -08:00
Richard Newman	6797a606b5	Preliminary work for vocabulary management. r=emily,nalexander Pre: export AttributeBuilder from mentat_db. Pre: fix module-level comment for tx/src/entities.rs. Pre: rename some `to_` conversions to `into_`. Pre: make AttributeBuilder::unique less verbose. Pre: split out a HasSchema trait to abstract over Schema. Pre: rename SchemaMap/schema_map to AttributeMap/attribute_map. Pre: TypedValue/NamespacedKeyword conversions. Pre: turn Unique and ValueType into TypedValue::Keyword. Pre: export IntoResult. Pre: export NamespacedKeyword from mentat_core. Pre: use intern_set in tx. Pre: add InternSet::len. Pre: comment gardening. Pre: remove inaccurate TODO from TxReport comment.	2018-01-23 08:25:32 -08:00
Thom	9740cafdbd	Automatically remove trailing whitespace from text files. (#527 ) r=rnewman This was done using the following shell script: ``` find . -type f -not -path "target" \ '(' -name '.rs' -o -name '.md' -o -name '.toml' ')' -print0 \| \ xargs -0 sed -i '' -E 's/[[:space:]]$//' ``` Which is admittedly imperfect, but manages to hit everything that was a problem in this repo.	2018-01-19 21:21:04 -06:00
Richard Newman	ebb77d59bc	Ensure that DateTime values are truncated to microsecond precision. (#522 ) r=emily	2018-01-18 09:06:23 -06:00
Richard Newman	75bcb76dd5	Be slightly more specific about mentat_core exporting Uuid.	2017-12-06 12:42:59 -08:00
Richard Newman	1f0c4e3107	Add common From coercions for TypedValue.	2017-12-06 12:29:11 -08:00
Richard Newman	df90c366af	Partial work from simple aggregates work (#497 ) r=nalexander * Pre: make FindQuery, FindSpec, and Element non-Clone. * Pre: make query translator return a Result. * Pre: make projection return a Result. * Pre: refactor query parser in preparation for parsing aggregates. * Pre: rename PredicateFn -> QueryFunction. * Pre: expose more about bound variables from CC. * Pre: move ValueTypeSet to core.	2017-11-30 15:02:07 -08:00
Richard Newman	c600152d78	Update some dependencies. (#492 ) r=etoop * Update some dependencies. * Update rusqlite to 0.12. * Update error-chain to a forked version that implements Sync. * Fix some compiler warnings. * Remove unused imports in tests. * Parse errors no longer naturally print with the expected symbol.	2017-11-21 16:24:08 +00:00
Richard Newman	c2ec1a6bdf	Pre: move Either to mentat_core::util.	2017-06-15 10:28:02 -07:00
Richard Newman	daca8def57	UUIDs and instants. Fixes #44 , #45 , #426 , #427 . (#438 ) r=nalexander * Pre: unused import in translate.rs. * Part 2: take a dependency on rusqlite for query arguments. * Part 1: flatten V2 schema into V1. Add UUID and URI. Bump expected ident and bootstrap datom count in tests. * Part 5: parse edn::Value::Uuid. * Part 3: extend ValueType and TypedValue to include Uuid. * Part 4: add Uuid to query arguments. * Part 6: extend db to support Uuid. * Part 8: add a tx-parser test for #f NaN and #uuid. * Part 7: parse and algebrize UUIDs in queries. * Part 1: parse #inst in EDN and throughout query engine. * Part 3: handle instants in db. * Part 2: instants never matches integers in queries. * Part 4: use DateTime for tx_instants. * Add a test for adding and querying UUIDs and instants. * Review comments.	2017-04-28 20:11:55 -07:00
Richard Newman	19fc7cddf1	[query] Widen `known_types` correctly in complex `or`. (#424 ) r=nalexander * Part 1: define ValueTypeSet. We're going to use this instead of `HashSet<ValueType>` so that we can clearly express the empty set and the set of all types, and also to encapsulate a switch to `EnumSet`." * Part 2: use ValueTypeSet. * Part 3: fix type expansion. * Part 4: add a test for type extraction from nested `or`. * Review comments. * Review comments: simplify ValueTypeSet.	2017-04-24 14:15:26 -07:00
Richard Newman	bffefe7e6b	Review comments for #418 .	2017-04-18 13:50:58 -07:00
Richard Newman	01af45ab3f	Pre: define Display for ValueType.	2017-04-18 13:19:50 -07:00
Richard Newman	98ac559894	Pre: allow initialization of a CC with an arbitrary counter value. Useful for testing.	2017-04-12 11:12:48 -07:00
Richard Newman	e984e02529	Pre: comment RcCounter.	2017-04-12 11:11:54 -07:00
Richard Newman	b117e2c463	Implement a cloneable shared counter. (#407 ) r=nalexander	2017-04-07 12:43:50 -07:00
Richard Newman	a5023c70cb	Use Rc for TypedValue, Variable, and query Ident keywords. (#395 ) r=nalexander Part 1, core: use Rc for String and Keyword. Part 2, query: use Rc for Variable. Part 3, sql: use Rc for args in SQLiteQueryBuilder. Part 4, query-algebrizer: use Rc. Part 5, db: use Rc. Part 6, query-parser: use Rc. Part 7, query-projector: use Rc. Part 8, query-translator: use Rc. Part 9, top level: use Rc. Part 10: intern Ident and IdentOrKeyword.	2017-04-02 21:38:36 -07:00
Richard Newman	8ae8466cf9	Make InternSet::intern accept Into<Rc<T>>. Add a test. r=nalexander	2017-03-31 09:59:58 -07:00
Richard Newman	460fdac252	Pre: add Variable::from_valid_name, TypedValue::{typed_string,typed_ns_keyword}.	2017-03-30 19:13:19 -07:00
Richard Newman	d2e6b767c6	Pre: add mentat_core::utils::{ResultEffect,OptionEffect}.	2017-03-30 19:06:28 -07:00
Emily Toop	8e6f37e709	#260 Convert Schema into edn::Value (#384 ) r=nalexander, r=rnewman * Part 1 - Create as_edn_value function. * Do not include defaults inside output. * Pretty-printed by default. Do we want to make that a flag? * Includes simple test just to make sure it works. * Part 2 - only include ident if available. * Part 3 - Remove spacing and newlines as unnecessary. * Update function to build edn::Value directly rather than parsing from string * Update test to actually test the functionality. * Address review comments ncalexan. * Rename `as_edn_value` to `to_edn_value`. * Move `db/src/values.rs` to `core/src/values.rs` so we can reference inside `core/src/ib.rs`. * Add `lazy-static` crate to core `Cargo.toml` * Expose `values` as a public module from `core`. * Update references to values in `db/src/bootstrap.rs` & `db/src/lib.rs`. * Add new static vars for `DB_FULLTEXT`, `DB_INDEX` & `DB_IS_COMPONENT`. * Use static vars exposed in `values` inside `to_edn_value`. * Remove `db/id` as key in attribute output and use `entid` as `db/ident` if no `ident` is found for that `entid`. * Update test to match new expected output. * Add doc comment for function * Address review comments ncalexan. * Update function docstring to give clearer description of function. * Do not all entid at all to output. * Clean up code fetching ident (make it rustier). * Address review comments rnewman. * Extract out to new `to_edn_value` functions code for creating `edn::Value`\'s for `ValueType` and `Attribute`. * Use `map()` to create schema edn value rather than a loop. * Address review comments rnewman. * pass cloned instance of ident to `Attribute::get_edn_value`. * update `use` import for `edn`. * remove unnecessary call when using ident as key on `associate_ident`. * Fixed bug whereby we didn't differentiate between `db.index/value` and `db.index/identity` when generating `edn::Value` * Add extra assert at the end to ensure we get the same output when we convert the same schema to edn multiple times * Move check for type of uniqueness to `match` statement. * Also use `iter` instead of `into_iter` when iterating schema map.	2017-03-30 11:08:36 +01:00
Nick Alexander	15b4195a6e	Schema alteration. Fixes #294 and #295 . (#370 ) r=rnewman * Pre: Don't retract :db/ident in test. Datomic (and eventually Mentat) don't allow to retract :db/ident in this way, so this runs afoul of future work to support mutating metadata. * Pre: s/VALUETYPE/VALUE_TYPE/. This is consistent with the capitalization (which is "valueType") and the other identifier. * Pre: Remove some single quotes from error output. * Part 1: Make materialized views be uniform [e a v value_type_tag]. This looks ahead to a time when we could support arbitrary user-defined materialized views. For now, the "idents" materialized view is those datoms of the form [e :db/ident :namespaced/keyword] and the "schema" materialized view is those datoms of the form [e a v] where a is in a particular set of attributes that will become clear in the following commits. This change is not backwards compatible, so I'm removing the open current (really, v2) test. It'll be re-instated when we get to https://github.com/mozilla/mentat/issues/194. * Pre: Map TypedValue::Ref to TypedValue::Keyword in debug output. * Part 3: Separate `schema_to_mutate` from the `schema` used to interpret. This is just to keep track of the expected changes during bootstrapping. I want bootstrap metadata mutations to flow through the same code path as metadata mutations during regular transactions; by differentiating the schema used for interpretation from the schema that will be updated I expect to be able to apply bootstrap metadata mutations to an empty schema and have things like materialized views created (using the regular code paths). This commit has been re-ordered for conceptual clarity, but it won't compile because it references the metadata module. It's possible to make it compile -- the functionality is there in the schema module -- but it's not worth the rebasing effort until after review (and possibly not even then, since we'll squash down to a single commit to land). * Part 2: Maintain entids separately from idents. In order to support historical idents, we need to distinguish the "current" map from entid -> ident from the "complete historical" map ident -> entid. This is what Datomic does; in Datomic, an ident is never retracted (although it can be replaced). This approach is an important part of allowing multiple consumers to share a schema fragment as it migrates forward. This fixes a limitation of the Clojure implementation, which did not handle historical idents across knowledge base close and re-open. The "entids" materialized view is naturally a slice of the "datoms" table. The "idents" materialized view is a slice of the "transactions" table. I hope that representing in this way, and casting the problem in this light, might generalize to future materialized views. * Pre: Add DiffSet. * Part 4: Collect mutations to a `Schema`. I haven't taken your review comment about consuming AttributeBuilder during each fluent function. If you read my response and still want this, I'm happy to do it in review. * Part 5: Handle :db/ident and :db.{install,alter}/attribute. This "loops" the committed datoms out of the SQL store and back through the metadata (schema, but in future also partition map) processor. The metadata processor updates the schema and produces a report of what changed; that report is then used to update the SQL store. That update includes: - the materialized views ("entids", "idents", and "schema"); - if needed, a subset of the datoms themselves (as flags change). I've left a TODO for handling attribute retraction in the cases that it makes sense. I expect that to be straight-forward. * Review comment: Rename DiffSet to AddRetractAlterSet. Also adds a little more commentary and a simple test. * Review comment: Use ToIdent trait. * Review comment: partially revert "Part 2: Maintain entids separately from idents." This reverts commit 23a91df9c35e14398f2ddbd1ba25315821e67401. Following our discussion, this removes the "entids" materialized view. The next commit will remove historical idents from the "idents" materialized view. * Post: Use custom Either rather than std::result::Result. This is not necessary, but it was suggested that we might be paying an overhead creating Err instances while using error_chain. That seems not to be the case, but this change shows that we don't actually use any of the Result helper methods, so there's no reason to overload Result. This change might avoid some future confusion, so I'm going to land it anyway. Signed-off-by: Nick Alexander <nalexander@mozilla.com> * Review comment: Don't preserve historical idents. * Review comment: More prepared statements when updating materialized views. * Post: Test altering :db/cardinality and :db/unique. These tests fail due to a Datomic limitation, namely that the marker flag :db.alter/attribute can only be asserted once for an attribute! That is, [:db.part/db :db.alter/attribute :attribute] will only be transacted at most once. Since older versions of Datomic required the :db.alter/attribute flag, I can only imagine they either never wrote :db.alter/attribute to the store, or they handled it specially. I'll need to remove the marker flag system from Mentat in order to address this fundamental limitation. * Post: Remove some more single quotes from error output. * Post: Add assert_transact! macro to unwrap safely. I was finding it very difficult to track unwrapping errors while making changes, due to an underlying Mac OS X symbolication issue that makes running tests with RUST_BACKTRACE=1 so slow that they all time out. * Post: Don't expect or recognize :db.{install,alter}/attribute. I had this all working... except we will never see a repeated `[:db.part/db :db.alter/attribute :attribute]` assertion in the store! That means my approach would let you alter an attribute at most one time. It's not worth hacking around this; it's better to just stop expecting (and recognizing) the marker flags. (We have all the data to distinguish the various cases that we need without the marker flags.) This brings Mentat in line with the thrust of newer Datomic versions, but isn't compatible with Datomic, because (if I understand correctly) Datomic automatically adds :db.{install,alter}/attribute assertions to transactions. I haven't purged the corresponding :db/ident and schema fragments just yet: - we might want them back - we might want them in order to upgrade v1 and v2 databases to the new on-disk layout we're fleshing out (v3?). * Post: Don't make :db/unique :db.unique/* imply :db/index true. This patch avoids a potential bug with the "schema" materialized view. If :db/unique :db.unique/value implies :db/index true, then what happens when you _retract_ :db.unique/value? I think Datomic defines this in some way, but I really want the "schema" materialized view to be a slice of "datoms" and not have these sort of ambiguities and persistent effects. Therefore, to ensure that we don't retract a schema characteristic and accidentally change more than we intended to, this patch stops having any schema characteristic imply any other schema characteristic(s). To achieve that, I added an Option<Unique::{Value,Identity}> type to Attribute; this helps with this patch, and also looks ahead to when we allow to retract :db/unique attributes. * Post: Allow to retract :db/ident. * Post: Include more details about invalid schema changes. The tests use strings, so they hide the chained errors which do in fact provide more detail. * Review comment: Fix outdated comment. * Review comment: s/_SET/_SQL_LIST/. * Review comment: Use a sub-select for checking cardinality. This might be faster in practice. * Review comment: Put `attribute::Unique` into its own namespace.	2017-03-20 13:18:59 -07:00
Richard Newman	bf38105fef	(#362 ) Part 4: handle unknown attributes by expanding type codes. r=nalexander Also, don't run any SQL at all if an algebrized query is known to return no results.	2017-03-08 17:44:27 -08:00
Richard Newman	7bcf311db9	Pre: move SQLValueType to core, because it's so central. Yes, this isn't tidy... but in order to be really tidy we'd need to split up db into parts that don't depend on a particular SQLite library.	2017-03-08 17:41:49 -08:00
Jordan Santell	6f67f8563b	TypedValue::Keyword now wraps a NamespacedKeyword rather than a String. Fixes #203 . r=nalexander" (#329 )	2017-02-17 14:07:57 -08:00
Jordan Santell	a59f9583ac	Store Idents as NamespacedKeywords, rather than Strings. Fixes #291 . (#300 ) r=ncalexander	2017-02-17 13:55:36 -08:00
Richard Newman	b165a0b2ad	Implement Schema::attribute_for_ident.	2017-02-15 16:01:22 -08:00
Nick Alexander	16e9740d8a	Implement upsert resolution algorithm. (#186 , #283 ). r=rnewman, f=jsantell * Pre: Implement batch [a v] pair lookup. * Pre: Add InternSet for sharing ref-counted handles to large values. * Pre: Derive more for Entity. * Pre: Return DB from creating; return TxReport from transact. I explicitly am not supporting opening existing databases yet, let alone upgrading databases from earlier versions. That can follow fast once basic transactions are supported. * Pre: Parse string temporary ID entities; remove ValueOrLookupRef. This adds TempId entities, but we can't disambiguate String temporary IDs from values without the use of the schema, so there's no new value branch. Similarly, we can't disambiguate lookup-ref values from two element list values without a schema, so we remove this entirely. We'll handle the ambiguity later in the transactor. * Persist partitions to SQL store; allocate transaction ID. (#186) * Post: Test upserting with vectors. This converts an existing test to EDN: `84a80f40f5/test/datomish/db_test.cljc (L193)`. * Implement tempid upsert resolution algorithm. (#184) * Post: Separate Tx out of DB. This is very preliminary, since we don't have a real connection type to manage transactions and their metadata yet. * Post: Comment on implementation choices in the transactor. * Review comment: Put long use lists on separate lines. * Review comment: Accept String: Borrow<S> instead of just String. * Review comment: Address nits.	2017-02-14 16:50:40 -08:00
Richard Newman	bfb62302cb	Add docstrings for Schema methods.	2017-02-13 19:39:47 -08:00
Richard Newman	a87e5a3ec7	Move Schema from mentat_db to mentat_core, improve API. (#290 ) * Move Schema from mentat_db to mentat_core. * Define SchemaMap in terms of Entid, not i64. * Add Schema::{is_attribute,identifies_attribute}. * Add pointer to #291. * Don't pass around 64-bit pointers to 64-bit integers.	2017-02-13 19:20:20 -08:00

1 2

55 commits