meilisearch

mirror of https://github.com/meilisearch/meilisearch.git synced 2025-11-22 04:36:32 +00:00

Author	SHA1	Message	Date
Clément Renault	ca4876fd10	Do not reindex when modifying unknown faceted field	2024-03-12 16:18:58 +01:00
meili-bors[bot]	ee3076d5ba	Merge #4462 4462: Divide threshold by ten r=dureuill a=ManyTheFish Change the facet incremental vs bulk indexing threshold to better fit our user needs, it might be changed in the future if we have more insights Co-authored-by: ManyTheFish <many@meilisearch.com>	2024-03-06 13:05:38 +00:00
meili-bors[bot]	ab1224bfa7	Merge #4458 4458: Replace logging timer by spans r=Kerollmops a=dureuill - Remove logging timer dependency. - Remplace last uses in search by spans Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2024-03-05 16:43:23 +00:00
meili-bors[bot]	eefc1c421e	Merge #4459 4459: Put a bound on OpenAI timeout r=dureuill a=dureuill # Pull Request ## Related issue Fixes #4460 ## What does this PR do? - Makes sure that the timeout of the openai embedder is limited to max 1min, rather than the prior 15min+ Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2024-03-05 15:18:51 +00:00
meili-bors[bot]	4d42a7af7c	Merge #4445 4445: Add subcommand to run benchmarks r=irevoire a=dureuill # Pull Request ## Related issue Not user-facing, no issue ## What does this PR do? - Adds a new `cargo xtask bench` subcommand that can run one or multiple workload files and report the results to a server - A workload file is a JSON file with a specific schema - Refactor our use of the `vergen` crate: - update to the beta `vergen-git2` crate - VERGEN_GIT_SEMVER_LIGHTWEIGHT => VERGEN_GIT_DESCRIBE - factor logic in a single `build-info` crate that is used both by meilisearch and xtask (prevents vergen variables from overriding themselves) - checked that defining the variables by hand when no git repo is available (docker build case) still works. - Add CI to run `cargo xtask bench` Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2024-03-05 14:03:57 +00:00
Louis Dureuil	7408db2a46	Meilisearch: fix date formatting	2024-03-05 14:56:48 +01:00
Louis Dureuil	663629a9d6	Remove unused build dependency from xtask Co-authored-by: Tamo <tamo@meilisearch.com>	2024-03-05 14:45:06 +01:00
Louis Dureuil	15c38dca78	Output RFC 3339 dates where we can Co-authored-by: Tamo <tamo@meilisearch.com>	2024-03-05 14:44:48 +01:00
Louis Dureuil	7ee20b0895	Refactor xtask bench	2024-03-05 14:42:06 +01:00
Louis Dureuil	0c216048b5	Cap timeout duration	2024-03-05 12:19:25 +01:00
Louis Dureuil	36d17110d8	openai: Handle BAD_GETAWAY, be more resilient to failure	2024-03-05 12:18:54 +01:00
meili-bors[bot]	bdd428c22e	Merge #4450 4450: Add the content type in the webhook + improve the test r=Kerollmops a=irevoire # Pull Request ## Related issue Fixes https://github.com/meilisearch/meilisearch/issues/4436 ## What does this PR do? - Specify the content type of the webhook - Ensure it’s the case in the test Co-authored-by: Tamo <tamo@meilisearch.com>	2024-03-05 10:36:53 +00:00
Tamo	b130917933	add the content type in the webhook + improve the test	2024-03-05 11:22:29 +01:00
Louis Dureuil	25f64ce7df	Replace logging timer by spans	2024-03-05 11:05:42 +01:00
Louis Dureuil	adcd848809	CI: Add bench workflows	2024-03-05 11:02:05 +01:00
Louis Dureuil	eee46b7537	Add first workloads	2024-03-05 10:13:11 +01:00
Louis Dureuil	55f60a3638	Update .gitignore - Ignore `/bench` directory for git purposes - Ignore benchmark DB	2024-03-05 10:12:52 +01:00
Louis Dureuil	c608b3f9b5	Factor vergen stuff to a build-info crate	2024-03-05 10:11:43 +01:00
Louis Dureuil	86ce843f3d	Add cargo xtask bench	2024-03-05 10:11:43 +01:00
Louis Dureuil	b11df7ec34	Meilisearch: fix some wrong spans	2024-03-05 10:11:43 +01:00
Louis Dureuil	6862caef64	Span Stats compute self-time	2024-03-05 10:11:43 +01:00
Louis Dureuil	f75c7ac979	Compile xtask in --release	2024-03-05 10:11:43 +01:00
ManyTheFish	eada6de261	Divide threshold by ten	2024-03-04 18:02:54 +01:00
meili-bors[bot]	f4a6261dea	Merge #4453 4453: Don't test on nightly r=dureuill a=dureuill # Pull Request ## Related issue Fixes #4441 better 😅 ## What does this PR do? - No longer run tests on nightly The motivation for this change is that we are now updating Rust at fixed points in time, and so no longer need nightly runs to ensure that a change won't get into stable and break our build at the worst possible moment. Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2024-02-29 14:41:59 +00:00
Louis Dureuil	9806a3e5f6	Don't test on nightly	2024-02-29 14:24:50 +01:00
meili-bors[bot]	a96b45dda7	Merge #4451 4451: Fix nightly build r=dureuill a=dureuill # Pull Request ## Related issue Fixes #4441 ## What does this PR do? - Change imports following https://github.com/rust-lang/rust/pull/117772 ## Note This one is going to be annoying a bit until the lint stabilizes: - We only get the warning on nightly, so we will discover them when it runs in the CI that uses the nightly compiler (not on regular PRs) - There's the case of `TryInto`/`TryFrom` traits. They have been added to the prelude in Rust edition 2021, so it means that `use`ing them is a warning on nightly for 2021 edition crates (most crates), but not `use`ing them is an error anywhere for 2018 Rust edition crates, such as `milli` Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2024-02-29 07:20:22 +00:00
Louis Dureuil	452a343a2b	Fix imports	2024-02-28 18:09:40 +01:00
meili-bors[bot]	b87485e80d	Merge #4433 4433: Enhance facet incremental r=Kerollmops a=ManyTheFish # Pull Request ## Related issue Fixes #4367 Fixes #4409 ## What does this PR do? - Add a test reproducing #4409 - Fix #4409 by removing a document from a level only if it is no more present in all the linked sub-level nodes - Optimize facet Incremental indexing by creating or deleting a complete level once per field id instead of for each facet value - Optimize facet Incremental indexing by doing the additions and the deletions in the same process instead of doing them separately Co-authored-by: ManyTheFish <many@meilisearch.com>	2024-02-28 15:28:46 +00:00
meili-bors[bot]	147a67dc82	Merge #4446 4446: Do not omit vectors when importing a dump r=irevoire a=dureuill # Pull Request ## Related issue Fixes #4447 ## What does this PR do? - Correctly populate the maps of embedders before starting the indexing operations, while importing a dump Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2024-02-27 09:11:00 +00:00
Louis Dureuil	716ffc07ee	Build the embedders when importing a dump	2024-02-26 22:15:57 +01:00
meili-bors[bot]	b005eb3289	Merge #4435 4435: Make update file deletion atomic r=Kerollmops a=irevoire # Pull Request ## Related issue Fixes https://github.com/meilisearch/meilisearch/issues/4432 Fixes https://github.com/meilisearch/meilisearch/issues/4438 by adding the logs the user asked ## What does this PR do? - Adds a bunch of logs to help debug this kind of issue in the future - Delete the update files AFTER committing the update in the `index-scheduler` (thus, if a restart happens, we are able to re-process the batch successfully) - Multi-thread the deletion of all update files. Co-authored-by: Tamo <tamo@meilisearch.com>	2024-02-26 17:54:40 +00:00
meili-bors[bot]	9e664d87eb	Merge #4443 4443: Add GPU analytics r=dureuill a=dureuill # Pull Request ## Related issue Adds analytics indicating whether Meilisearch was compiled with the `milli/cuda` feature. Cc `@macraig` Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2024-02-26 17:13:45 +00:00
meili-bors[bot]	6dcb5219a0	Merge #4442 4442: Send custom task r=ManyTheFish a=irevoire This PR has already been merged on main but was supposed to be merged on `release-v1.7.0` thus we need to merge it a second time; sorry 😓 ### This PR implements the necessary parameters for the High Availability Introduce a new CLI flag called `--experimental-replication-parameters` that changes a few behaviors in the engine: - [The auto-deletion of tasks is disabled](https://specs.meilisearch.com/specifications/text/0060-tasks-api.html#_2-technical-details) - Upon registering a task, you can choose its task ID by sending a new header: `TaskId: 456645`. It must be a valid number, which must be superior to the last task id ever seen. - Add the ability to « dry-register » a task. That means meilisearch will answer to you with a valid task ID like everything went well, but won’t actually write anything in the database. To do that, you need to use the `DryRun: true` header. - Specification’s here: https://github.com/meilisearch/specifications/pull/266 Co-authored-by: Tamo <tamo@meilisearch.com>	2024-02-26 15:20:16 +00:00
ManyTheFish	5e83bac448	Fix PR comments	2024-02-26 15:40:15 +01:00
Tamo	0562818c2a	fix and remove the file-store hack of /dev/null	2024-02-26 13:59:41 +01:00
Tamo	a478392b7a	create a test with the dry-run parameter enabled	2024-02-26 13:59:41 +01:00
Tamo	bbf3fb88ca	rename the cli parameter	2024-02-26 13:59:40 +01:00
Tamo	60510e037b	update the discussion link	2024-02-26 13:58:04 +01:00
Tamo	36c27a18a1	implement the dry run ha parameter	2024-02-26 13:58:04 +01:00
Tamo	1eb1c043b5	disable the auto deletion of tasks when the ha mode is enabled	2024-02-26 13:58:04 +01:00
Tamo	507739bd98	add an experimental cli parameter to allow specifying your task id	2024-02-26 13:58:03 +01:00
Tamo	eb25b07390	let you specify your task id	2024-02-26 13:56:31 +01:00
Tamo	066a7a3cde	takes only one read transaction per thread	2024-02-26 10:43:04 +01:00
Louis Dureuil	55796406c5	Add GPU analytics	2024-02-26 10:41:47 +01:00
Tamo	91cdd502f8	When processing tasks, make the update file deletion atomic	2024-02-22 14:56:22 +01:00
ManyTheFish	a493a50825	Fix clippy	2024-02-22 14:53:33 +01:00
ManyTheFish	9d1f489a37	Fix facet incremental indexing	2024-02-21 18:42:16 +01:00
ManyTheFish	865b415b3f	Add test rerpoducing bug	2024-02-15 16:00:48 +01:00
meili-bors[bot]	5ee6aaddc4	Merge #4418 4418: Output logs to stderr r=dureuill a=irevoire Output the logs to `stderr` instead of `stdout`. This was introduced in the `v1.7.0-rc.0` and is a bug; logs should always be outputted to stderr. Fix #4419 Co-authored-by: Tamo <tamo@meilisearch.com>	2024-02-15 14:31:37 +00:00
Tamo	4148d391b8	move logs to stderr	2024-02-15 15:24:16 +01:00
meili-bors[bot]	88c6165e20	Merge #4410 4410: Implement the experimental log mode cli flag and log level updates at runtime r=dureuill a=irevoire # Pull Request This PR fixes two issues at once because they’re highly correlated in the codebase. ## Related issue Fixes https://github.com/meilisearch/meilisearch/issues/4415 Fixes https://github.com/meilisearch/meilisearch/issues/4413 ## What does this PR do? - It makes the fmt logger configurable to output json or human-readable logs (like we already do today) - It moves the fmt logger under a `reload` layer so we can update its targets at runtime - Add the possibility to stream logs in the json mode - Adds an analytics for the new CLI flag Co-authored-by: Tamo <tamo@meilisearch.com>	2024-02-15 10:01:06 +00:00
Tamo	d097431113	Update meilisearch/src/option.rs Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2024-02-15 10:58:43 +01:00
Tamo	1f8af81ba9	update the log mode discussion link	2024-02-15 10:32:48 +01:00
Tamo	5d3bad4120	Update meilisearch/src/option.rs Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2024-02-15 10:31:23 +01:00
meili-bors[bot]	d34692e30b	Merge #4365 4365: Update charabia r=dureuill a=ManyTheFish Update Charabia v0.8.7, - Add Vietnamese Normalization (Ð and Đ into d) Fixes #4357 Charabia versions: - https://github.com/meilisearch/charabia/releases/tag/v0.8.6 - https://github.com/meilisearch/charabia/releases/tag/v0.8.7 Co-authored-by: ManyTheFish <many@meilisearch.com>	2024-02-14 16:57:25 +00:00
Tamo	a081da0d90	add support for the json format in the stream route	2024-02-14 15:34:39 +01:00
ManyTheFish	78e04520fc	Update charabia version	2024-02-14 15:16:16 +01:00
meili-bors[bot]	72c1674a31	Merge #4350 4350: Make several indexing optimizations r=Kerollmops a=ManyTheFish # Summary Implement several enhancements to reduce the indexing time. # Steps - Compute the indexing chunk size dynamically based on the available threads and the data size - Remove the merging step before the writing step and merge at the writing time - Remove append function - Make Facet search indexing incremental # Running Indexing process ## `main` Each type of data is written after a merging phase: ![Capture d’écran 2024-01-23 à 10 18 08](https://github.com/meilisearch/meilisearch/assets/6482087/6203c3ce-407c-46b4-8b83-04282da1bb16) > Highlighted parts are the writings ## `remove-merging-phase-from-indexing` When the extraction of a chunk is finished, the data is written: ![Capture d’écran 2024-01-23 à 10 18 18](https://github.com/meilisearch/meilisearch/assets/6482087/ab1307b4-d0a9-42ac-abbb-fdeb27ddf0d4) > Highlighted parts are the writings ## Related This PR removes the appending writes on several indexing parts, which may fix https://github.com/meilisearch/meilisearch/issues/4300. However, all of the appending writes are not removed. There are 2 remaining calls that could trigger this bug: - When [putting embedders in the settings](`b6fc181993/milli/src/update/settings.rs (L996)`) - when [bulk indexing the facets](`b6fc181993/milli/src/update/facet/bulk.rs (L150)`) Co-authored-by: ManyTheFish <many@meilisearch.com> Co-authored-by: Louis Dureuil <louis@meilisearch.com> Co-authored-by: Many the fish <many@meilisearch.com>	2024-02-14 14:12:48 +00:00
ManyTheFish	03bb6372af	Change is_batchable_with by mergeable_with	2024-02-14 11:50:22 +01:00
ManyTheFish	3beda8833d	Fix and add logs	2024-02-14 11:46:30 +01:00
Tamo	3b6544db6d	Implement the experimental log mode cli flag	2024-02-13 18:09:15 +01:00
ManyTheFish	55e942cd45	buggy	2024-02-13 15:26:30 +01:00
ManyTheFish	48026aa75c	fix PR comments	2024-02-13 15:19:01 +01:00
Many the fish	e5e811e2c9	Update milli/src/update/index_documents/extract/mod.rs Co-authored-by: Clément Renault <clement@meilisearch.com>	2024-02-13 14:22:21 +01:00
Many the fish	55de96f74e	Update milli/src/update/facet/mod.rs Co-authored-by: Clément Renault <clement@meilisearch.com>	2024-02-13 14:22:10 +01:00
meili-bors[bot]	15dafde21d	Merge #4401 4401: Update version for the next release (v1.7.0) in Cargo.toml r=irevoire a=meili-bot ⚠️ This PR is automatically generated. Check the new version is the expected one and Cargo.lock has been updated before merging. Co-authored-by: irevoire <irevoire@users.noreply.github.com>	2024-02-12 10:17:10 +00:00
irevoire	290f6d15e7	Update version for the next release (v1.7.0) in Cargo.toml	2024-02-12 10:15:00 +00:00
ManyTheFish	39c83cb3d9	fix clippy	2024-02-12 09:12:54 +01:00
Louis Dureuil	7efb1cae11	yield in loop when the channel is not disconnected	2024-02-12 09:12:54 +01:00
Louis Dureuil	7877788510	fix logs	2024-02-12 09:12:54 +01:00
ManyTheFish	be1b054b05	Compute chunk size based on the input data size ant the number of indexing threads	2024-02-08 17:28:37 +01:00
meili-bors[bot]	023c2d755f	Merge #4391 4391: Tracing r=dureuill a=irevoire # Pull Request - [ ] Hide the parameters of the process batch - [x] Make actix-web trace every call on every route - [x] Remove all `env_logger`/`logs` dependencies - [x] Be able to enable or disable the memory measurement using the `/logs` route parameters See the following product discussion: https://github.com/orgs/meilisearch/discussions/721 Supersedes https://github.com/meilisearch/meilisearch/pull/4338 ## Related issue Fixes https://github.com/meilisearch/meilisearch/issues/4317 ## What does this PR do? Update the format of the logs from: ``` [2024-02-06T14:54:11Z INFO actix_server::builder] starting 10 workers ``` to ``` 2024-02-06T13:58:14.710803Z INFO actix_server::builder: 200: starting 10 workers ``` First, run meilisearch with the route enabled via the feature flag: - `cargo run --experimental-enable-logs-route` - Or at runtime by sending the following payload: ``` curl \ -X PATCH 'http://localhost:7700/experimental-features/' \ -H 'Content-Type: application/json' \ --data-binary '{ "logsRoute": true }' ``` Then gather data from meilisearch by calling for example: ``` curl \ -X POST http://localhost:7700/logs \ -H 'Content-Type: application/json' \ --data-binary '{ "mode": "fmt", "target": "milli=trace" }' ``` Once your operation is over, tell meilisearch to stop the route: ``` curl \ -X DELETE http://localhost:7700/logs ``` ---- In the case you’re profiling code, you will be interested by the next command that converts the output of the route to a format that the firefox profiler can understand. ```bash cargo run --release --bin trace-to-firefox -- 2024-01-17_17:07:55-indexing-trace.json ``` Then go to https://profiler.firefox.com and load it. Note that we can also share the profiles using the https://share.firefox.dev website. Co-authored-by: Louis Dureuil <louis@meilisearch.com> Co-authored-by: Clément Renault <clement@meilisearch.com> Co-authored-by: Tamo <tamo@meilisearch.com>	2024-02-08 14:16:56 +00:00
Louis Dureuil	407ad753ed	rust fmt	2024-02-08 15:11:42 +01:00
Tamo	285aa15d2f	make the mode camelCase instead of lowercase	2024-02-08 15:04:06 +01:00
Tamo	bf43a3f60a	fix typo	2024-02-08 15:04:06 +01:00
Tamo	2c88131bb1	rename the fmt mode to human	2024-02-08 15:04:06 +01:00
Tamo	35aa9d5904	fix an error message	2024-02-08 15:04:06 +01:00
Tamo	cfb3e6b51f	update the actix-web trace	2024-02-08 15:04:06 +01:00
Tamo	1502382316	use debug instead of debug_span	2024-02-08 15:04:06 +01:00
Louis Dureuil	ef994d84d0	Change error messages and fix tests	2024-02-08 15:04:06 +01:00
Louis Dureuil	1b74010e9e	Remove "with_line_numbers"	2024-02-08 15:04:06 +01:00
Tamo	08af0e690c	Structures a bunch of logs	2024-02-08 15:04:06 +01:00
Louis Dureuil	d71b77f18b	Add panic hook to log panics	2024-02-08 15:04:06 +01:00
Louis Dureuil	c443ed7e3f	delete inner .gitignore	2024-02-08 15:04:06 +01:00
Louis Dureuil	db722d201a	Write entries into database downgraded to trace level	2024-02-08 15:04:05 +01:00
Louis Dureuil	91eb67e981	logs route: make memory profiling toggling usable	2024-02-08 15:04:05 +01:00
Louis Dureuil	902d700a24	Tracing trace: toggle the profiling of memory at runtime	2024-02-08 15:04:05 +01:00
Tamo	f70a615ed9	update the github discussion links	2024-02-08 15:04:05 +01:00
Tamo	7ff722b72e	get rids of the log dependencies everywhere	2024-02-08 15:04:05 +01:00
Tamo	bcf7909bba	add a profile_memory parameter disabled by default	2024-02-08 15:04:05 +01:00
Tamo	ceb211c515	move the /logs route to the /logs/stream route	2024-02-08 15:04:05 +01:00
Clément Renault	f3c34d5b8c	Simplify MemoryStats fetching	2024-02-08 15:04:05 +01:00
Tamo	4de2db6786	add back the actix-web logs	2024-02-08 15:04:05 +01:00
Louis Dureuil	661baa716b	logs route profile mode: don't barf bytes if the buffer is not empty	2024-02-08 15:04:05 +01:00
Clément Renault	02dcaf07db	Replace the procfs by libproc	2024-02-08 15:04:05 +01:00
Louis Dureuil	d78ada07b5	spanstats: change field names	2024-02-08 15:04:05 +01:00
Louis Dureuil	bc097d90cb	tracing-trace: Spanstats deserializable + public fields	2024-02-08 15:04:05 +01:00
Clément Renault	b393823f36	Replace stats_alloc with procfs	2024-02-08 15:04:05 +01:00
Tamo	e773dfa9ba	get rids of log in milli and add logs for the bucket sort	2024-02-08 15:04:05 +01:00
Tamo	f158e96fe7	fix the auth	2024-02-08 15:04:05 +01:00
Tamo	e23ec4886d	fix the tests and add tests on the experimental features	2024-02-08 15:04:03 +01:00
Tamo	7793ba67a4	hide the route logs behind a feature flag	2024-02-08 15:03:33 +01:00
Tamo	80774148fd	handle and tests errors	2024-02-08 15:03:33 +01:00
Tamo	bf5cea8b10	add a test	2024-02-08 15:03:33 +01:00
Louis Dureuil	38e1c40f38	meilisearch: logs route disconnects in profile mode	2024-02-08 15:03:33 +01:00
Louis Dureuil	afc0585c1c	meilisearch: don't spawn a report everytime Meilisearch starts	2024-02-08 15:03:33 +01:00
Louis Dureuil	0e7a411d4d	tracing-trace: introduce TraceWriter, trace now only exposes the channel	2024-02-08 15:03:33 +01:00
Louis Dureuil	0f327f2821	tracing-trace: implement Error on Error	2024-02-08 15:03:33 +01:00
Tamo	77254765e8	get rids of env loggegr and fix the tests	2024-02-08 15:03:33 +01:00
Tamo	ce6e6ec2c5	stops profiling in a file by default	2024-02-08 15:03:32 +01:00
Louis Dureuil	91a8f74763	Add cancel log route	2024-02-08 15:03:32 +01:00
Tamo	abaa72e2bf	start handling reloads with profiling	2024-02-08 15:03:32 +01:00
Tamo	3c3a258a22	start exposing the profiling layer	2024-02-08 15:03:32 +01:00
Louis Dureuil	73e66d5a97	Add dummy log when calling tasks	2024-02-08 15:03:32 +01:00
Louis Dureuil	b8da117b9c	Simplify stream implementation	2024-02-08 15:03:32 +01:00
Louis Dureuil	5e52107474	better than before???	2024-02-08 15:03:32 +01:00
Tamo	bcf1c4dae5	make it compile and runtime error	2024-02-08 15:03:32 +01:00
Tamo	50f84d43f5	init commit	2024-02-08 15:03:32 +01:00
Tamo	f76cc0806e	WIP: first draft at introducing a new log route	2024-02-08 15:03:32 +01:00
Tamo	2f1abd2c03	nelson is not used anymore	2024-02-08 15:03:32 +01:00
Louis Dureuil	dedc91e2cf	use json lines	2024-02-08 15:03:32 +01:00
Louis Dureuil	a61d8c59ff	Add span stats processor	2024-02-08 15:03:32 +01:00
Louis Dureuil	6e23040464	Use with tokio channel in Meilisearch	2024-02-08 15:03:32 +01:00
Louis Dureuil	8febbf64ce	Switch to tokio channel	2024-02-08 15:03:32 +01:00
Louis Dureuil	b141c82a04	Support Events in trace layer	2024-02-08 15:03:32 +01:00
Louis Dureuil	cc79cd0b04	Switch to a single view indicating current usage	2024-02-08 15:03:32 +01:00
Louis Dureuil	256538ccb9	Refactor memory handling and add markers	2024-02-08 15:03:31 +01:00
Clément Renault	ca8990394e	Remove the stats_alloc from the default features	2024-02-08 15:03:31 +01:00
Clément Renault	83fb2949c3	Give the allocator to the tracer when necessary	2024-02-08 15:03:31 +01:00
Louis Dureuil	6cf703387d	Format the bytes as human readable bytes Uses the same `byte_unit` version as `meilisearch`	2024-02-08 15:03:31 +01:00
Clément Renault	771861599b	Logging the memory usage over time	2024-02-08 15:03:31 +01:00
Louis Dureuil	7e47cea0c4	Add tracing to Meilisearch	2024-02-08 15:03:31 +01:00
Louis Dureuil	5d7061682e	Add tracing to milli	2024-02-08 15:03:31 +01:00
Louis Dureuil	02e6c8a440	Add tracing to index-scheduler	2024-02-08 15:03:31 +01:00
Louis Dureuil	89401d097b	Add tracing-trace	2024-02-08 15:03:30 +01:00
meili-bors[bot]	72ebac1fbb	Merge #4388 4388: Cap the maximum memory of the grenad sorters r=curquiza a=Kerollmops This PR clamps the memory usage of the grenad sorters to a reasonable maximum. Grenad sorters are opened on multiple threads at a time. This can result in higher memory usage than expected, even though it shouldn't consume more than the memory available. Fixes #4152. Co-authored-by: Clément Renault <clement@meilisearch.com>	2024-02-08 13:19:28 +00:00
meili-bors[bot]	a616a1d37b	Merge #4389 4389: Stabilize scoreDetails r=dureuill a=dureuill # Pull Request ## Related issue Fixes #4359 ## What does this PR do? ### User standpoint - Users no longer need to enable the `scoreDetails` experimental feature to use `showRankingScoreDetails` in search queries. - ⚠️ Breaking change: sending an object containing the key `"scoreDetails"` to the `/experimental-features` route is now an error. However, importing a dump of a database where that feature was enabled completes successfully. ### Implementation standpoint - remove `scoreDetails` from the experimental features - remove check on the experimental feature `scoreDetails` before accepting `showRankingScoreDetails` - remove `scoreDetails` from the accepted fields in the `/experimental-features` route - fix tests accordingly ## Manual tests 1. exported a dump with the `scoreDetails` feature enabled on `main` - tried to import the dump after the changes in this PR - the dump imported successfully 2. tried to make a search with `showRankingScoreDetails: true` - the ranking score details are displayed - an automated test case also exists and passes 3. tried to enable the `scoreDetails` in `/experimental-features` - get error message ``` Unknown field `scoreDetails`: expected one of `vectorStore`, `metrics`, `exportPuffinReports` ``` Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2024-02-08 10:40:00 +00:00
meili-bors[bot]	3e120619fa	Merge #4375 4375: Feat: add new OpenAI models and ability to override dimensions r=dureuill a=Gosti # Pull Request Fixes #4394 ## Related discussion https://github.com/orgs/meilisearch/discussions/677#discussioncomment-8306384 ## What does this PR do? - Add text-embedding-3-small - Add text-embedding-3-large - Add optional dimensions parameter for both new models ## Note As the dimensions option is not available for text-embedding-ada-002 I've added a manual check to prevent, but I feel it could be implemented in a more idiomatic rust ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Gosti <gostitsog@gmail.com> Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2024-02-07 16:20:15 +00:00
Louis Dureuil	a1caac9bfb	Correct distribution shifts for new models	2024-02-07 15:09:16 +01:00
Louis Dureuil	88d03c56ab	Don't accept dimensions of 0 (ever) or dimensions greater than the default dimensions of the model	2024-02-07 11:52:09 +01:00
Louis Dureuil	32ee05ccef	Fix default dimensions for models	2024-02-07 11:52:09 +01:00
Louis Dureuil	74c180267e	pass dimensions only when defined	2024-02-07 11:52:08 +01:00
Louis Dureuil	517f5332d6	Allow actually passing `dimensions` for OpenAI source -> make sure the settings change is rejected or the settings task fails when the specified model doesn't support overriding `dimensions` and the passed `dimensions` differs from the model's default dimensions.	2024-02-07 11:51:44 +01:00
Louis Dureuil	9ac5750096	Retrieve the overriden dimensions from the configuration when fetching settings	2024-02-07 11:51:44 +01:00
Louis Dureuil	7ae4013478	Make sure the overriden dimensions are always used when embedding	2024-02-07 11:51:44 +01:00
Gosti	fb705116a6	feat: add new models and ability to override dimensions	2024-02-07 11:51:42 +01:00
Clément Renault	053306c0e7	Try with 500MiB	2024-02-07 11:24:43 +01:00
meili-bors[bot]	84235a63df	Merge #4360 4360: fix readme broken links r=curquiza a=Elliot67 # Pull Request ## What does this PR do? - fix some links in the readme ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Co-authored-by: Elliot Lintz <45725915+Elliot67@users.noreply.github.com> Co-authored-by: gui machiavelli <hey@guimachiavelli.com>	2024-02-06 16:00:16 +00:00
gui machiavelli	29f8300ac7	Update README.md	2024-02-06 16:49:29 +01:00
Louis Dureuil	05edd85d75	Stabilize scoreDetails	2024-02-06 11:15:19 +01:00
Clément Renault	9eeb75d501	Clamp the max memory of the grenad sorters to a reasonable maximum	2024-02-06 10:47:04 +01:00
meili-bors[bot]	4792651462	Merge #4384 4384: Bump peter-evans/repository-dispatch from 2 to 3 r=curquiza a=dependabot[bot] Bumps [peter-evans/repository-dispatch](https://github.com/peter-evans/repository-dispatch) from 2 to 3. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/peter-evans/repository-dispatch/releases">peter-evans/repository-dispatch's releases</a>.</em></p> <blockquote> <h2>Repository Dispatch v3.0.0</h2> <p>⚙️ Updated runtime to Node.js 20</p> <ul> <li>The action now requires a minimum version of <a href="https://github.com/actions/runner/releases/tag/v2.308.0">v2.308.0</a> for the Actions runner. Update self-hosted runners to v2.308.0 or later to ensure compatibility.</li> </ul> <h2>What's Changed</h2> <ul> <li>Bump prettier to fix deps by <a href="https://github.com/peter-evans"><code>`@peter-evans</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/255">peter-evans/repository-dispatch#255</a></li> <li>build(deps-dev): bump <code>`@types/node</code>` from 18.17.12 to 18.17.14 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/257">peter-evans/repository-dispatch#257</a></li> <li>build(deps-dev): bump <code>`@vercel/ncc</code>` from 0.36.1 to 0.38.0 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/258">peter-evans/repository-dispatch#258</a></li> <li>build(deps): bump actions/checkout from 3 to 4 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/259">peter-evans/repository-dispatch#259</a></li> <li>build(deps-dev): bump <code>`@types/node</code>` from 18.17.14 to 18.17.16 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/261">peter-evans/repository-dispatch#261</a></li> <li>build(deps): bump <code>`@actions/core</code>` from 1.10.0 to 1.10.1 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/262">peter-evans/repository-dispatch#262</a></li> <li>build(deps-dev): bump jest-circus from 29.6.4 to 29.7.0 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/263">peter-evans/repository-dispatch#263</a></li> <li>build(deps-dev): bump eslint from 8.48.0 to 8.49.0 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/264">peter-evans/repository-dispatch#264</a></li> <li>Update distribution by <a href="https://github.com/actions-bot"><code>`@actions-bot</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/265">peter-evans/repository-dispatch#265</a></li> <li>build(deps-dev): bump <code>`@types/node</code>` from 18.17.16 to 18.17.18 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/266">peter-evans/repository-dispatch#266</a></li> <li>build(deps-dev): bump eslint-plugin-github from 4.10.0 to 4.10.1 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/267">peter-evans/repository-dispatch#267</a></li> <li>build(deps-dev): bump <code>`@types/node</code>` from 18.17.18 to 18.18.0 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/268">peter-evans/repository-dispatch#268</a></li> <li>build(deps-dev): bump eslint from 8.49.0 to 8.50.0 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/269">peter-evans/repository-dispatch#269</a></li> <li>build(deps-dev): bump <code>`@types/node</code>` from 18.18.0 to 18.18.3 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/271">peter-evans/repository-dispatch#271</a></li> <li>build(deps-dev): bump eslint-plugin-prettier from 5.0.0 to 5.0.1 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/275">peter-evans/repository-dispatch#275</a></li> <li>build(deps-dev): bump <code>`@types/node</code>` from 18.18.3 to 18.18.5 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/274">peter-evans/repository-dispatch#274</a></li> <li>build(deps-dev): bump eslint from 8.50.0 to 8.51.0 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/276">peter-evans/repository-dispatch#276</a></li> <li>build(deps-dev): bump <code>`@babel/traverse</code>` from 7.16.3 to 7.23.2 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/278">peter-evans/repository-dispatch#278</a></li> <li>build(deps-dev): bump <code>`@types/node</code>` from 18.18.5 to 18.18.6 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/279">peter-evans/repository-dispatch#279</a></li> <li>build(deps-dev): bump <code>`@vercel/ncc</code>` from 0.38.0 to 0.38.1 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/280">peter-evans/repository-dispatch#280</a></li> <li>build(deps-dev): bump eslint from 8.51.0 to 8.52.0 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/281">peter-evans/repository-dispatch#281</a></li> <li>build(deps-dev): bump <code>`@types/node</code>` from 18.18.6 to 18.18.7 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/282">peter-evans/repository-dispatch#282</a></li> <li>build(deps): bump actions/setup-node from 3 to 4 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/283">peter-evans/repository-dispatch#283</a></li> <li>build(deps-dev): bump <code>`@types/node</code>` from 18.18.7 to 18.18.8 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/284">peter-evans/repository-dispatch#284</a></li> <li>build(deps-dev): bump <code>`@types/node</code>` from 18.18.8 to 18.18.9 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/285">peter-evans/repository-dispatch#285</a></li> <li>build(deps-dev): bump eslint from 8.52.0 to 8.53.0 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/286">peter-evans/repository-dispatch#286</a></li> <li>build(deps-dev): bump prettier from 3.0.3 to 3.1.0 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/287">peter-evans/repository-dispatch#287</a></li> <li>build(deps-dev): bump eslint from 8.53.0 to 8.54.0 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/289">peter-evans/repository-dispatch#289</a></li> <li>build(deps-dev): bump <code>`@types/node</code>` from 18.18.9 to 18.18.13 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/290">peter-evans/repository-dispatch#290</a></li> <li>build(deps-dev): bump <code>`@types/node</code>` from 18.18.13 to 18.19.0 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/291">peter-evans/repository-dispatch#291</a></li> <li>build(deps-dev): bump <code>`@types/node</code>` from 18.19.0 to 18.19.3 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/292">peter-evans/repository-dispatch#292</a></li> <li>build(deps-dev): bump eslint from 8.54.0 to 8.55.0 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/293">peter-evans/repository-dispatch#293</a></li> <li>build(deps-dev): bump prettier from 3.1.0 to 3.1.1 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/296">peter-evans/repository-dispatch#296</a></li> <li>build(deps): bump actions/upload-artifact from 3 to 4 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/295">peter-evans/repository-dispatch#295</a></li> <li>build(deps-dev): bump eslint from 8.55.0 to 8.56.0 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/297">peter-evans/repository-dispatch#297</a></li> <li>build(deps-dev): bump eslint-plugin-prettier from 5.0.1 to 5.1.1 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/298">peter-evans/repository-dispatch#298</a></li> <li>build(deps-dev): bump eslint-plugin-prettier from 5.1.1 to 5.1.2 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/299">peter-evans/repository-dispatch#299</a></li> <li>build(deps-dev): bump <code>`@types/node</code>` from 18.19.3 to 18.19.4 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/300">peter-evans/repository-dispatch#300</a></li> <li>build(deps-dev): bump eslint-plugin-prettier from 5.1.2 to 5.1.3 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/301">peter-evans/repository-dispatch#301</a></li> <li>build(deps-dev): bump <code>`@types/node</code>` from 18.19.4 to 18.19.6 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/302">peter-evans/repository-dispatch#302</a></li> <li>build(deps-dev): bump prettier from 3.1.1 to 3.2.4 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/303">peter-evans/repository-dispatch#303</a></li> <li>build(deps-dev): bump <code>`@types/node</code>` from 18.19.6 to 18.19.8 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/304">peter-evans/repository-dispatch#304</a></li> <li>feat: update runtime to node 20 by <a href="https://github.com/peter-evans"><code>`@peter-evans</code></a>` in <a href="https://redirect.github.com/peter-evans/repository-dispatch/pull/305">peter-evans/repository-dispatch#305</a></li> </ul> <!-- raw HTML omitted --> </blockquote> <p>... (truncated)</p> </details> <details> <summary>Commits</summary> <ul> <li><a href="`ff45666b94`"><code>ff45666</code></a> feat: update runtime to node 20 (<a href="https://redirect.github.com/peter-evans/repository-dispatch/issues/305">#305</a>)</li> <li><a href="`a4a90276d0`"><code>a4a9027</code></a> build(deps-dev): bump <code>`@types/node</code>` from 18.19.6 to 18.19.8 (<a href="https://redirect.github.com/peter-evans/repository-dispatch/issues/304">#304</a>)</li> <li><a href="`2605253283`"><code>2605253</code></a> build(deps-dev): bump prettier from 3.1.1 to 3.2.4 (<a href="https://redirect.github.com/peter-evans/repository-dispatch/issues/303">#303</a>)</li> <li><a href="`ab3258eeef`"><code>ab3258e</code></a> build(deps-dev): bump <code>`@types/node</code>` from 18.19.4 to 18.19.6 (<a href="https://redirect.github.com/peter-evans/repository-dispatch/issues/302">#302</a>)</li> <li><a href="`240bc73193`"><code>240bc73</code></a> build(deps-dev): bump eslint-plugin-prettier from 5.1.2 to 5.1.3 (<a href="https://redirect.github.com/peter-evans/repository-dispatch/issues/301">#301</a>)</li> <li><a href="`8aa15c54a0`"><code>8aa15c5</code></a> build(deps-dev): bump <code>`@types/node</code>` from 18.19.3 to 18.19.4 (<a href="https://redirect.github.com/peter-evans/repository-dispatch/issues/300">#300</a>)</li> <li><a href="`22aa07cf23`"><code>22aa07c</code></a> build(deps-dev): bump eslint-plugin-prettier from 5.1.1 to 5.1.2 (<a href="https://redirect.github.com/peter-evans/repository-dispatch/issues/299">#299</a>)</li> <li><a href="`ba0298574b`"><code>ba02985</code></a> build(deps-dev): bump eslint-plugin-prettier from 5.0.1 to 5.1.1 (<a href="https://redirect.github.com/peter-evans/repository-dispatch/issues/298">#298</a>)</li> <li><a href="`accfd7b5bf`"><code>accfd7b</code></a> build(deps-dev): bump eslint from 8.55.0 to 8.56.0 (<a href="https://redirect.github.com/peter-evans/repository-dispatch/issues/297">#297</a>)</li> <li><a href="`3c7d964ae9`"><code>3c7d964</code></a> build(deps): bump actions/upload-artifact from 3 to 4 (<a href="https://redirect.github.com/peter-evans/repository-dispatch/issues/295">#295</a>)</li> <li>Additional commits viewable in <a href="https://github.com/peter-evans/repository-dispatch/compare/v2...v3">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=peter-evans/repository-dispatch&package-manager=github_actions&previous-version=2&new-version=3)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) You can trigger a rebase of this PR by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` show <dependency name> ignore conditions` will show all of the ignore conditions of the specified dependency - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) </details> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2024-02-05 14:13:35 +00:00
dependabot[bot]	58c3501b54	Bump peter-evans/repository-dispatch from 2 to 3 Bumps [peter-evans/repository-dispatch](https://github.com/peter-evans/repository-dispatch) from 2 to 3. - [Release notes](https://github.com/peter-evans/repository-dispatch/releases) - [Commits](https://github.com/peter-evans/repository-dispatch/compare/v2...v3) --- updated-dependencies: - dependency-name: peter-evans/repository-dispatch dependency-type: direct:production update-type: version-update:semver-major ... Signed-off-by: dependabot[bot] <support@github.com>	2024-02-01 17:05:50 +00:00
meili-bors[bot]	ff76d8f21a	Merge #4382 4382: Bring back changes from `release-v1.6.1` into `main` r=curquiza a=dureuill Bring back changes from release-v1.6.1 into main Supersedes https://github.com/meilisearch/meilisearch/pull/4380 and #4381 Third time's the charm Co-authored-by: curquiza <curquiza@users.noreply.github.com> Co-authored-by: Louis Dureuil <louis@meilisearch.com> Co-authored-by: Tamo <tamo@meilisearch.com> Co-authored-by: Morgane Dubus <30866152+mdubus@users.noreply.github.com>	2024-02-01 11:16:31 +00:00
Louis Dureuil	698ea5139d	Update Cargo.lock	2024-02-01 10:40:23 +01:00
Morgane Dubus	880e790bff	Update Cargo.toml	2024-02-01 10:33:27 +01:00
Louis Dureuil	fbf5f2a392	Don't use a runtime in extract_embedder, use it only for OpenAI	2024-02-01 10:33:27 +01:00
Louis Dureuil	1555870088	Truncate HuggingFace vectors that are too long	2024-02-01 10:33:27 +01:00
Tamo	9f8f3105d5	make clippy happy	2024-02-01 10:33:27 +01:00
Tamo	318843aacd	add a bunch of tests and fix the error message when adding the geosearch as filterable/sortable while there is malformed documents in the DB	2024-02-01 10:33:27 +01:00
Louis Dureuil	6d111139b5	Add test	2024-02-01 10:33:27 +01:00
Louis Dureuil	dff2707471	Use MatchingWords from keyword search instead of the one from vector search	2024-02-01 10:33:27 +01:00
curquiza	c57f7f7379	Update version for the next release (v1.6.1) in Cargo.toml	2024-02-01 10:33:26 +01:00
meili-bors[bot]	b968616a99	Merge #4364 4364: Revert "Remove panic on the geosearch" r=curquiza a=irevoire After more thought about it, we want to fix this bug in a patch release instead of `main`. I revert this PR for now, but the fix will still land on `main` once we bring back the change of the `v1.6.1` on `main`. Reverts meilisearch/meilisearch#4337 Co-authored-by: Tamo <irevoire@protonmail.ch>	2024-01-25 18:01:08 +00:00
Tamo	c1bf33a112	Revert "Remove panic on the geosearch"	2024-01-25 18:51:19 +01:00
Elliot Lintz	ddc2b7129a	fix readme broken links	2024-01-24 22:50:18 +01:00
meili-bors[bot]	b6fc181993	Merge #4304 4304: Add CUDA GPU support for Hugging Face embedders r=Kerollmops a=dureuill Adds a "cuda" feature to `milli`. Compiling with this feature requires that the CUDA support library be installed (see "with CUDA support" paragraph in https://huggingface.github.io/candle/guide/installation.html), and adds CUDA support to the `huggingFace` embedder. To enable GPU support, users will need to: 1. Have a compatible NVidia GPU under Linux 2. Follow [the guide](https://huggingface.github.io/candle/guide/installation.html) to install the CUDA dependencies 3. Compile Meilisearch with the `cuda` feature: `cargo build --release --features cuda` # Impact Enabling the CUDA feature allows to use an available GPU to compute embeddings with a `huggingFace` embedder. On an AWS Graviton 2, this yields a x3 - x5 improvement on indexing time. # Technical details - I had to change the CI so that the cuda feature is not included in the `Tests all features` workflow - To achieve that, I had to add a binary following the `cargo xtask` design pattern, to list all features excepted the cuda one. - I then changed the workflow accordingly (renamed to "Tests almost all features" 😉) - A test run of the new feature was done on a temporary version of this PR that had it enabled for PRs: [See the results here](https://github.com/meilisearch/meilisearch/actions/runs/7461331929/job/20301216732) Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2024-01-22 13:55:04 +00:00
meili-bors[bot]	388fce9e46	Merge #4345 4345: Bump h2 from 0.3.20 to 0.3.24 r=curquiza a=dependabot[bot] Bumps [h2](https://github.com/hyperium/h2) from 0.3.20 to 0.3.24. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/hyperium/h2/releases">h2's releases</a>.</em></p> <blockquote> <h2>v0.3.24</h2> <h2>Fixed</h2> <ul> <li>Limit error resets for misbehaving connections.</li> </ul> <h2>v0.3.23</h2> <h2>What's Changed</h2> <ul> <li>cherry-pick fix: streams awaiting capacity lockout in <a href="https://redirect.github.com/hyperium/h2/pull/734">hyperium/h2#734</a></li> </ul> <h2>v0.3.22</h2> <h2>What's Changed</h2> <ul> <li>Add <code>header_table_size(usize)</code> option to client and server builders.</li> <li>Improve throughput when vectored IO is not available.</li> <li>Update indexmap to 2.</li> </ul> <h2>New Contributors</h2> <ul> <li><a href="https://github.com/tottoto"><code>`@tottoto</code></a>` made their first contribution in <a href="https://redirect.github.com/hyperium/h2/pull/714">hyperium/h2#714</a></li> <li><a href="https://github.com/xiaoyawei"><code>`@xiaoyawei</code></a>` made their first contribution in <a href="https://redirect.github.com/hyperium/h2/pull/712">hyperium/h2#712</a></li> <li><a href="https://github.com/Protryon"><code>`@Protryon</code></a>` made their first contribution in <a href="https://redirect.github.com/hyperium/h2/pull/719">hyperium/h2#719</a></li> <li><a href="https://github.com/4JX"><code>`@4JX</code></a>` made their first contribution in <a href="https://redirect.github.com/hyperium/h2/pull/638">hyperium/h2#638</a></li> <li><a href="https://github.com/vuittont60"><code>`@vuittont60</code></a>` made their first contribution in <a href="https://redirect.github.com/hyperium/h2/pull/724">hyperium/h2#724</a></li> </ul> <h2>v0.3.21</h2> <h2>What's Changed</h2> <ul> <li>Fix opening of new streams over peer's max concurrent limit.</li> <li>Fix <code>RecvStream</code> to return data even if it has received a <code>CANCEL</code> stream error.</li> <li>Update MSRV to 1.63.</li> </ul> <h2>New Contributors</h2> <ul> <li><a href="https://github.com/DDtKey"><code>`@DDtKey</code></a>` made their first contribution in <a href="https://redirect.github.com/hyperium/h2/pull/703">hyperium/h2#703</a></li> <li><a href="https://github.com/jwilm"><code>`@jwilm</code></a>` made their first contribution in <a href="https://redirect.github.com/hyperium/h2/pull/707">hyperium/h2#707</a></li> </ul> </blockquote> </details> <details> <summary>Changelog</summary> <p><em>Sourced from <a href="https://github.com/hyperium/h2/blob/v0.3.24/CHANGELOG.md">h2's changelog</a>.</em></p> <blockquote> <h1>0.3.24 (January 17, 2024)</h1> <ul> <li>Limit error resets for misbehaving connections.</li> </ul> <h1>0.3.23 (January 10, 2024)</h1> <ul> <li>Backport fix from 0.4.1 for stream capacity assignment.</li> </ul> <h1>0.3.22 (November 15, 2023)</h1> <ul> <li>Add <code>header_table_size(usize)</code> option to client and server builders.</li> <li>Improve throughput when vectored IO is not available.</li> <li>Update indexmap to 2.</li> </ul> <h1>0.3.21 (August 21, 2023)</h1> <ul> <li>Fix opening of new streams over peer's max concurrent limit.</li> <li>Fix <code>RecvStream</code> to return data even if it has received a <code>CANCEL</code> stream error.</li> <li>Update MSRV to 1.63.</li> </ul> </blockquote> </details> <details> <summary>Commits</summary> <ul> <li><a href="`7243ab5854`"><code>7243ab5</code></a> Prepare v0.3.24</li> <li><a href="`d919cd6fd8`"><code>d919cd6</code></a> streams: limit error resets for misbehaving connections</li> <li><a href="`a7eb14a487`"><code>a7eb14a</code></a> v0.3.23</li> <li><a href="`b668c7fbe2`"><code>b668c7f</code></a> fix: streams awaiting capacity lockout (<a href="https://redirect.github.com/hyperium/h2/issues/730">#730</a>) (<a href="https://redirect.github.com/hyperium/h2/issues/734">#734</a>)</li> <li><a href="`0f412d8b9c`"><code>0f412d8</code></a> v0.3.22</li> <li><a href="`c7ca62f69b`"><code>c7ca62f</code></a> docs: fix typos (<a href="https://redirect.github.com/hyperium/h2/issues/724">#724</a>)</li> <li><a href="`ef743ecb22`"><code>ef743ec</code></a> Add a setter for header_table_size (<a href="https://redirect.github.com/hyperium/h2/issues/638">#638</a>)</li> <li><a href="`56651e6e51`"><code>56651e6</code></a> fix lint about unused import</li> <li><a href="`4aa7b16342`"><code>4aa7b16</code></a> Fix documentation for max_send_buffer_size (<a href="https://redirect.github.com/hyperium/h2/issues/718">#718</a>)</li> <li><a href="`d03c54a80d`"><code>d03c54a</code></a> chore(dependencies): update tracing minimal version to 0.1.35</li> <li>Additional commits viewable in <a href="https://github.com/hyperium/h2/compare/v0.3.20...v0.3.24">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=h2&package-manager=cargo&previous-version=0.3.20&new-version=0.3.24)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) Dependabot will resolve any conflicts with this PR as long as you don't alter it yourself. You can also trigger a rebase manually by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` show <dependency name> ignore conditions` will show all of the ignore conditions of the specified dependency - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) You can disable automated security fix PRs for this repo from the [Security Alerts page](https://github.com/meilisearch/meilisearch/network/alerts). </details> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2024-01-22 11:53:51 +00:00
Louis Dureuil	d35fe43fd5	Update lock file	2024-01-22 10:49:17 +01:00
Louis Dureuil	f692021bfc	Implement PR comments	2024-01-22 10:25:56 +01:00
Louis Dureuil	1b90778bf5	Change CI	2024-01-22 10:25:56 +01:00
Louis Dureuil	66ae81a909	Make it so binary can be used with `cargo xtask`	2024-01-22 10:25:56 +01:00
Louis Dureuil	4aa4a15dc9	Add to Cargo.lock	2024-01-22 10:25:54 +01:00
Louis Dureuil	4b4e8ea2a4	Add binary to list features	2024-01-22 10:25:16 +01:00
Louis Dureuil	84f49d76cd	Add cuda feature	2024-01-22 10:25:16 +01:00
meili-bors[bot]	afb0e8eab9	Merge #4325 4325: Add Setting API reminder in issue template r=ManyTheFish a=ManyTheFish When adding a new setting, several important points can be easily forgotten. This PR adds a small reminder list of some of these points in the issue template. Co-authored-by: Many the fish <many@meilisearch.com>	2024-01-22 09:02:27 +00:00
dependabot[bot]	b5b2333a05	Bump h2 from 0.3.20 to 0.3.24 Bumps [h2](https://github.com/hyperium/h2) from 0.3.20 to 0.3.24. - [Release notes](https://github.com/hyperium/h2/releases) - [Changelog](https://github.com/hyperium/h2/blob/v0.3.24/CHANGELOG.md) - [Commits](https://github.com/hyperium/h2/compare/v0.3.20...v0.3.24) --- updated-dependencies: - dependency-name: h2 dependency-type: indirect ... Signed-off-by: dependabot[bot] <support@github.com>	2024-01-19 16:20:22 +00:00
Many the fish	40fa0b4df6	Update .github/ISSUE_TEMPLATE/sprint_issue.md	2024-01-18 11:17:29 +01:00
Many the fish	ab4d614599	Update .github/ISSUE_TEMPLATE/sprint_issue.md Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2024-01-18 10:28:30 +01:00
meili-bors[bot]	262b20fdba	Merge #4330 4330: Add job variable to grafana dashboard r=irevoire a=capJavert # Pull Request ## Related issue Fixes https://github.com/orgs/meilisearch/discussions/625#discussioncomment-8143282 ## What does this PR do? "meilisearch" as [job_name](https://prometheus.io/docs/prometheus/latest/configuration/configuration/#job_name) was hardcoded in the dashboard config so if user sets anything but "meilisearch" as job_name on prometheus side the dashboard does not work. With this change dasboard will auto load the values from data source (much like instance variable) and show the correct data. This now also adds support for multiple meilisearch jobs in single dashboard. See: https://github.com/orgs/meilisearch/discussions/625#discussioncomment-8143282 ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: capJavert <ante@kickass.website>	2024-01-17 15:48:24 +00:00
meili-bors[bot]	9020606c45	Merge #4337 4337: Remove panic on the geosearch r=ManyTheFish a=irevoire # Pull Request ## Related issue Fixes #4333 ## What does this PR do? - Add tests for the enrich pipeline on malformed documents with `null` value - Reproduce the issue when updating the settings while there is malformed documents in the DB - Fix the bug Co-authored-by: Tamo <tamo@meilisearch.com>	2024-01-17 15:09:46 +00:00
Tamo	0887186ecf	make clippy happy	2024-01-17 16:07:10 +01:00
Tamo	7d190d8078	add a bunch of tests and fix the error message when adding the geosearch as filterable/sortable while there is malformed documents in the DB	2024-01-17 15:51:52 +01:00
meili-bors[bot]	3b8a9597e2	Merge #4332 4332: Update the dependencies r=irevoire a=Kerollmops This PR upgrades the dependencies and fixes #4287. - ~We keep arroy at the current commit. We will release and use the latest version published when possible~ - We also updated arroy to 0.2.0. - I rolled back the version of rustls has too many breaking changes. - I had to keep HTTP to 0.2.11 due to actix-cors. Co-authored-by: Clément Renault <clement@meilisearch.com>	2024-01-17 13:42:02 +00:00
Clément Renault	f275554982	Make sure we override the default Rust version	2024-01-16 18:10:30 +01:00
Clément Renault	d997ea1f01	Make Clippy happy	2024-01-16 17:10:48 +01:00
Clément Renault	50e1d34c66	Rollback http to 0.2.11	2024-01-16 16:57:33 +01:00
Clément Renault	406531c991	Fix sysinfo	2024-01-16 16:49:51 +01:00
Clément Renault	01e2c3d6bb	Bump arroy to v0.2.0	2024-01-16 16:45:55 +01:00
Clément Renault	cfaa522d68	Bump the Rust version to 1.75.0	2024-01-16 16:36:54 +01:00
Clément Renault	0c8d1644a6	Rollback rustls to 0.20.9	2024-01-16 15:55:16 +01:00
Clément Renault	5e0268d40e	Fix the sysinfo errors	2024-01-16 15:43:03 +01:00
Clément Renault	9f9ad4cc05	Fix Clippy warnings	2024-01-16 15:27:24 +01:00
Clément Renault	3ee7682fa7	Fix some integer comparisons	2024-01-16 15:22:23 +01:00
Clément Renault	7f125bfb12	Update incompatible dependencies	2024-01-16 15:15:54 +01:00
Clément Renault	5869ca7716	Upgrade all compatible dependencies	2024-01-16 15:05:03 +01:00
meili-bors[bot]	7a89abd2a0	Merge #4263 4263: Bump rustls-webpki from 0.101.3 to 0.101.7 r=irevoire a=dependabot[bot] Bumps [rustls-webpki](https://github.com/rustls/webpki) from 0.101.3 to 0.101.7. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/rustls/webpki/releases">rustls-webpki's releases</a>.</em></p> <blockquote> <h2>0.101.7</h2> <ul> <li>Upgrades <code>ring</code> to 0.17, and <code>untrusted</code> to 0.9. Note: since <code>untrusted</code> appears in the <code>Error</code> API this may be a breaking change for applications using two <code>untrusted</code> versions.</li> </ul> <h2>What's Changed</h2> <ul> <li>Simplify tests for DER errors by <a href="https://github.com/djc"><code>`@djc</code></a>` in <a href="https://redirect.github.com/rustls/webpki/pull/193">rustls/webpki#193</a></li> <li>Upgrade to ring 0.17, untrusted 0.9 by <a href="https://github.com/djc"><code>`@djc</code></a>` in <a href="https://redirect.github.com/rustls/webpki/pull/193">rustls/webpki#193</a></li> <li>Bump MSRV to 1.61 by <a href="https://github.com/djc"><code>`@djc</code></a>` in <a href="https://redirect.github.com/rustls/webpki/pull/193">rustls/webpki#193</a></li> <li>Upgrade to rcgen 0.11.3 by <a href="https://github.com/cpu"><code>`@cpu</code></a>` in <a href="https://redirect.github.com/rustls/webpki/pull/189">rustls/webpki#189</a>, <a href="https://redirect.github.com/rustls/webpki/pull/195">rustls/webpki#195</a></li> <li>v0.101.7 preparation by <a href="https://github.com/cpu"><code>`@cpu</code></a>` in <a href="https://redirect.github.com/rustls/webpki/pull/199">rustls/webpki#199</a></li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/rustls/webpki/compare/v/0.101.6...v/0.101.7">https://github.com/rustls/webpki/compare/v/0.101.6...v/0.101.7</a></p> <h2>0.101.6</h2> <ul> <li>The <code>CertificateRevocationList</code> trait's <code>verify_signature</code> <code>Budget</code> argument was removed. This was a semver incompatible change mistakenly introduced in v0.101.5.</li> </ul> <h2>What's Changed</h2> <ul> <li>crl: rm Budget from verify_signature fn by <a href="https://github.com/cpu"><code>`@cpu</code></a>` in <a href="https://redirect.github.com/rustls/webpki/pull/187">rustls/webpki#187</a></li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/rustls/webpki/compare/v/0.101.5...v/0.101.6">https://github.com/rustls/webpki/compare/v/0.101.5...v/0.101.6</a></p> <h2>0.101.5</h2> <ul> <li>Path building complexity is now limited to a maximum budget of path finding operations, avoiding exponential processing time when encountering certificate chains containing many certificates with the same subject/issuer distinguished name but different subject public key information.</li> <li>Name constraints evaluation is now limited to a maximum number of comparison operations, avoiding exponential processing time when encountering certificate chains containing many name constraints and subject alternate names.</li> <li>Subject common names are no longer parsed for name iteration, or applying name constraints. Webpki only uses Subject Alternate Names when validating certificates, and the common name handling was buggy, producing <code>Error::BadDer</code> when iterating certificates with printable string subject common names, or omitted common names encoded as an empty sequence.</li> </ul> <h2>What's Changed</h2> <p>The following PRs were backported to the rel-0.101 branch in <a href="https://redirect.github.com/rustls/webpki/issues/170">#170</a>:</p> <ul> <li>Further limits on expensive path building (<a href="https://redirect.github.com/rustls/webpki/issues/163">#163</a>)</li> <li>Budget tweaks (<a href="https://redirect.github.com/rustls/webpki/issues/164">#164</a>)</li> <li>Bound name constraint comparisons (<a href="https://redirect.github.com/rustls/webpki/issues/165">#165</a>)</li> <li>Remove subject common name parsing (<a href="https://redirect.github.com/rustls/webpki/issues/169">#169</a>, thanks to <a href="https://github.com/hawkw"><code>`@hawkw</code></a>)</li>` <li>Correct handling of fatal errors (<a href="https://redirect.github.com/rustls/webpki/issues/168">#168</a>)</li> </ul> <p>Thanks to all who have contributed, on behalf of the rustls team (<a href="https://github.com/ctz"><code>`@ctz</code></a>,` <a href="https://github.com/cpu"><code>`@cpu</code></a>` and <a href="https://github.com/djc"><code>`@djc</code></a>)!</p>` <h2>0.101.4</h2> <h2>Release notes</h2> <ul> <li>certificate path building and verification is now capped at 100 signature validation operations to avoid the risk of CPU usage denial-of-service attack when validating crafted certificate chains producing quadratic runtime. This risk affected both clients, as well as servers that verified client certificates.</li> </ul> <h2>What's Changed</h2> <ul> <li>v0.101.4 prep by <a href="https://github.com/cpu"><code>`@cpu</code></a>` in <a href="https://redirect.github.com/rustls/webpki/pull/153">rustls/webpki#153</a></li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/rustls/webpki/compare/v/0.101.3...v/0.101.4">https://github.com/rustls/webpki/compare/v/0.101.3...v/0.101.4</a></p> </blockquote> </details> <details> <summary>Commits</summary> <ul> <li><a href="`ee5aab1dff`"><code>ee5aab1</code></a> Cargo: v0.101.6 -> v0.101.7</li> <li><a href="`4f721a901f`"><code>4f721a9</code></a> Upgrade to rcgen 0.11.3</li> <li><a href="`3be3625584`"><code>3be3625</code></a> Bump MSRV to 1.61</li> <li><a href="`bb7c7f47ab`"><code>bb7c7f4</code></a> Upgrade to ring 0.17, untrusted 0.9</li> <li><a href="`2eeb2920cf`"><code>2eeb292</code></a> Simplify tests for DER errors</li> <li><a href="`7956538ee7`"><code>7956538</code></a> Cargo: v0.101.5 -> v0.101.6</li> <li><a href="`7f8208ec06`"><code>7f8208e</code></a> crl: rm <code>Budget</code> from <code>verify_signature</code> fn</li> <li><a href="`7cb6c646a0`"><code>7cb6c64</code></a> Cargo: bump version 0.101.4 -> 0.101.5</li> <li><a href="`2dd2a06016`"><code>2dd2a06</code></a> verify_cert: use enum for build chain error</li> <li><a href="`c255d61a6a`"><code>c255d61</code></a> verify_cert: correct handling of fatal errors</li> <li>Additional commits viewable in <a href="https://github.com/rustls/webpki/compare/v/0.101.3...v/0.101.7">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=rustls-webpki&package-manager=cargo&previous-version=0.101.3&new-version=0.101.7)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) Dependabot will resolve any conflicts with this PR as long as you don't alter it yourself. You can also trigger a rebase manually by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` show <dependency name> ignore conditions` will show all of the ignore conditions of the specified dependency - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) You can disable automated security fix PRs for this repo from the [Security Alerts page](https://github.com/meilisearch/meilisearch/network/alerts). </details> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2024-01-16 13:55:49 +00:00
Clément Renault	d9d0419845	Update the dependencies	2024-01-16 14:38:48 +01:00
capJavert	5dc8d9e9bf	feat: add job variable to dashboard meilisearch job_name was hardcoded in the dashboard config so if user sets anything but meilisearch as job_name on prometheus side the dashboard does not work see: https://github.com/orgs/meilisearch/discussions/625#discussioncomment-8143282	2024-01-16 12:44:37 +01:00
Many the fish	9e12a91afb	Update .github/ISSUE_TEMPLATE/sprint_issue.md	2024-01-16 11:04:50 +01:00
meili-bors[bot]	8e016fbfeb	Merge #4319 4319: Update README r=curquiza a=codesmith-emmy # Pull Request ## Related issue Fixes #<issue_number> ## What does this PR do? - ... ## PR checklist Please check if your PR fulfills the following requirements: - [ ] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [ ] Have you read the contributing guidelines? - [ ] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: emmanuel <154705254+codesmith-emmy@users.noreply.github.com>	2024-01-15 18:41:14 +00:00
meili-bors[bot]	1ccde9bf0b	Merge #4316 4316: Autobatch the task deletions r=curquiza a=irevoire # Pull Request ## Related issue Fix part of https://github.com/meilisearch/meilisearch-support/issues/69 Fix #4315 ## What does this PR do? - Autobatch the task deletions Co-authored-by: Tamo <tamo@meilisearch.com>	2024-01-15 17:54:50 +00:00
meili-bors[bot]	34e814f400	Merge #4327 4327: Bring back changes from `release-v1.6.0` to `main` r=dureuill a=curquiza Co-authored-by: Paul Sanders <psanders1@gmail.com> Co-authored-by: meili-bors[bot] <89034592+meili-bors[bot]@users.noreply.github.com> Co-authored-by: Louis Dureuil <louis.dureuil@xinra.net> Co-authored-by: Tamo <tamo@meilisearch.com> Co-authored-by: Clément Renault <clement@meilisearch.com> Co-authored-by: Louis Dureuil <louis@meilisearch.com> Co-authored-by: Morgane Dubus <30866152+mdubus@users.noreply.github.com>	2024-01-15 16:52:05 +00:00
Many the fish	857cd09285	Add Setting API reminder in issue template When adding a new setting, there are several important points that can be easily forgotten. This PR adds a small reminder list of some of these points.	2024-01-15 11:19:13 +01:00
meili-bors[bot]	a6fa0b97ec	Merge #4318 4318: Hide embedders r=ManyTheFish a=dureuill Hides `embedders` when it is an empty dictionary. Manual tests: - getting settings with empty embedders: not displayed - getting settings with non-empty embedders: displayed like before - dump with empty embedders: can be imported - dump with non-empty embedders: can be imported Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2024-01-15 09:37:31 +00:00
emmanuel	552127021f	Update	2024-01-12 16:03:23 +01:00
Louis Dureuil	38abfec611	Fix tests	2024-01-11 21:35:30 +01:00
Louis Dureuil	84a5c304fc	Don't display the embedders setting when it is an empty dict	2024-01-11 21:35:06 +01:00
meili-bors[bot]	e93d36d5b9	Merge #4313 4313: Fix document formatting performances r=Kerollmops a=ManyTheFish reduce the formatted option list to the attributes that should be formatted, instead of all the attributes to display. The time to compute the `format` list scales with the number of fields to format; cumulated with `map_leaf_values` that iterates over all the nested fields, it gives a quadratic complexity: `d*f` where `d` is the total number of fields to display and `f` is the total number of fields to format. Co-authored-by: ManyTheFish <many@meilisearch.com>	2024-01-11 14:19:44 +00:00
ManyTheFish	95f8e21533	fix typos	2024-01-11 15:07:08 +01:00
Tamo	b4d7d80ad9	autobatch the task deletions	2024-01-11 14:58:07 +01:00
meili-bors[bot]	68f197624e	Merge #4314 4314: Fix proximity precision telemetry r=Kerollmops a=ManyTheFish The proximity precision telemetry was partially missing in the global setting route. This PR adds the missing field and return the default value when the value is not set. Co-authored-by: ManyTheFish <many@meilisearch.com>	2024-01-11 13:50:03 +00:00
ManyTheFish	b79b03d4e2	Fix proximity precision telemetry	2024-01-11 13:24:26 +01:00
ManyTheFish	86270e6878	Transform fields contained into _format into strings	2024-01-11 12:44:56 +01:00
ManyTheFish	81b6128b29	Update tests	2024-01-11 12:28:32 +01:00
ManyTheFish	5f5a486895	Reduce formatting time	2024-01-11 11:36:41 +01:00
ManyTheFish	5f4fc6c955	Add timer logs	2024-01-11 09:44:16 +01:00
meili-bors[bot]	1f5e8fc072	Merge #4311 4311: Limit the number of values returned by the facet search r=dureuill a=Kerollmops This PR fixes a bug where the number of values per facet returned by the `indexes/{index}/facet-search` route was not tacking the `faceting.maxValuePerFacet` setting. It also adds a test. Co-authored-by: Clément Renault <clement@meilisearch.com>	2024-01-10 16:04:06 +00:00
Clément Renault	3f3462ab62	Limit the number of values returned by the facet search	2024-01-10 16:54:08 +01:00
meili-bors[bot]	93363b0201	Merge #4308 4308: Fix hang on `/indexes` and `/stats` routes r=Kerollmops a=dureuill # Pull Request ## Related issue Fixes #4218 ## Context - A previous fix added a field to the `IndexScheduler` to memorize the `currently_updating_index`, so that accessing it through the search would return the handle without trying to open it. This resolved a hang on the search, but #4218 reported further hangs on the `/indexes` and `/stats` routes - These routes were shunting the `IndexScheduler` and using internal `IndexMapper` logic to access the indexes, again trying to reopen the updating index. ## What does this PR do? - Moves the logic relative to the `currently_updating_index` from the `IndexScheduler` to the `IndexMapper`, so that any index request to the `IndexMapper` can benefit from it. ## Test 1. Follow reproducer from #4218 2. Before this PR, notice a hang on `/stats` and `/indexes`, but not on `/indexes/<updating_index>/search` 3. After this PR, notice no hang on either of `/stats`, `/indexes` or `/indexes/<updating_index>/search` Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2024-01-10 10:46:20 +00:00
Louis Dureuil	97bb1ff9e2	Move `currently_updating_index` to IndexMapper	2024-01-09 15:37:27 +01:00
meili-bors[bot]	5ee1378856	Merge #4303 4303: Display default value when proximityPrecision is not set r=dureuill a=ManyTheFish # Pull Request ## Related Issue: #4187 Spec change requests: https://github.com/meilisearch/specifications/pull/261#discussion_r1441725272 ## What does this PR do? - Display default value when proximityPrecision is not set instead of Null Co-authored-by: ManyTheFish <many@meilisearch.com>	2024-01-08 14:29:57 +00:00
ManyTheFish	e27b850b09	move the default display strategy on setting getter function	2024-01-08 14:03:47 +01:00
ManyTheFish	f75f22e026	Display default value when proximityPrecision is not set	2024-01-08 11:09:37 +01:00
meili-bors[bot]	6203f4acef	Merge #4296 4296: Fix single element search r=irevoire a=dureuill # Pull Request Before this PR, indexing a single vector in a single document would result in the vector not being found by the vector search. This PR adds a test case for this condition, and resolves it by bumping arroy to a version containing the fix. # Test case Output of the test before and after this PR: ```diff diff --git a/meilisearch/tests/search/hybrid.rs b/meilisearch/tests/search/hybrid.rs index 2cd4b83e7..79819cab2 100644 --- a/meilisearch/tests/search/hybrid.rs on release-v1.6.0 +++ b/meilisearch/tests/search/hybrid.rs on fix-single-element-search `@@` -171,5 +171,5 `@@` async fn single_document() { .await; snapshot!(code, `@"200` OK"); - snapshot!(response["hits"][0], `@r###"{"title":"Shazam!","desc":"a` Captain Marvel ersatz","id":"1","_vectors":{"default":[1.0,3.0]},"_rankingScore":0.0}"###); + snapshot!(response["hits"][0], `@r###"{"title":"Shazam!","desc":"a` Captain Marvel ersatz","id":"1","_vectors":{"default":[1.0,3.0]},"_rankingScore":1.0,"_semanticScore":1.0}"###); } ``` Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2024-01-03 15:01:43 +00:00
Louis Dureuil	12edc2c20a	Update arroy to a fixed version	2024-01-03 15:59:37 +01:00
Louis Dureuil	94b9f3b310	Add test	2024-01-03 15:56:20 +01:00
meili-bors[bot]	5204c0b60b	Merge #4297 4297: Update license for 2024 r=curquiza a=meili-bot _This PR is auto-generated._ Co-authored-by: meili-bot <74670311+meili-bot@users.noreply.github.com>	2024-01-03 13:54:19 +00:00
meili-bot	e73cd692db	Update LICENSE	2024-01-03 14:32:41 +01:00
meili-bors[bot]	29b453346b	Merge #4293 4293: Update SDK test dependencies r=curquiza a=curquiza Replace dependabot updates The changes are really un-impactful for the engine team velocity because is about a CI - that does not run during release deployment - that does not run to merge a PR It's only a weekly scheduled CI to check the breaking we introduced in the integrations. I updated the dependencies based on what we do on the integration CIs For example for dart, I looked at what we have in the [Dart CI](`63fd758882/.github/workflows/tests.yml (L16-L54)`) and I updated our CI in this repo accordingly. I did the same for each repository. This ensures we test the same things. Co-authored-by: curquiza <clementine@meilisearch.com>	2024-01-03 13:26:50 +00:00
meili-bors[bot]	c4bb435374	Merge #4295 4295: fix compilation warnings on main r=curquiza a=irevoire # Pull Request ## Related issue Fixes https://github.com/meilisearch/meilisearch/issues/4292 ## What does this PR do? - Removed unused imports #4294 fixes the issue for the release v1.6 Co-authored-by: Tamo <tamo@meilisearch.com>	2024-01-02 15:33:06 +00:00
meili-bors[bot]	da99a04eb3	Merge #4294 4294: fix compilation warnings for release v1.6 r=curquiza a=irevoire # Pull Request ## Related issue Fixes #4292 ## What does this PR do? - Removed unused imports #4295 fixes the issue no main Co-authored-by: Tamo <tamo@meilisearch.com>	2024-01-02 15:00:40 +00:00
Tamo	54ae6951eb	fix warning	2024-01-02 15:19:30 +01:00
Tamo	2bcff2ea46	fix warning	2024-01-02 15:19:00 +01:00
curquiza	1275e72e0b	Update SDK test dependencies	2024-01-02 09:59:46 +01:00
meili-bors[bot]	658ec6e0a4	Merge #4279 4279: Check experimental feature on setting update query rather than in the task. r=ManyTheFish a=dureuill Improve the UX by checking for the vector store feature and returning an error synchronously when sending a setting update, rather than in the indexing task. Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-12-22 11:36:12 +00:00
meili-bors[bot]	43e822e802	Merge #4238 4238: Task queue webhook r=dureuill a=irevoire # Prototype `prototype-task-queue-webhook-1` The prototype is available through Docker by using the following command: ```bash docker run -p 7700:7700 -v $(pwd)/meili_data:/meili_data getmeili/meilisearch:prototype-task-queue-webhook-1 ``` # Pull Request Implements the task queue webhook. ## Related issue Fixes https://github.com/meilisearch/meilisearch/issues/4236 ## What does this PR do? - Provide a new cli and env var for the webhook, respectively called `--task-webhook-url` and `MEILI_TASK_WEBHOOK_URL` - Also supports sending the requests with a custom `Authorization` header by specifying the optional `--task-webhook-authorization-header` CLI parameter or `MEILI_TASK_WEBHOOK_AUTHORIZATION_HEADER` env variable. - Throw an error if the specified URL is invalid - Every time a batch is processed, send all the finished tasks into the webhook with our public `TaskView` type as a JSON Line GZIPed body. - Add one test. ## PR checklist ### Before becoming ready to review - [x] Add a test - [x] Compress the data we send - [x] Chunk and stream the data we send - [x] Remove the unwrap in the index-scheduler when sending the data fails - [x] The analytics are missing ### Before merging - [x] Release a prototype Co-authored-by: Tamo <tamo@meilisearch.com> Co-authored-by: Clément Renault <clement@meilisearch.com>	2023-12-21 14:43:46 +00:00
Louis Dureuil	ee54d3171e	Check experimental feature at query time	2023-12-21 15:26:12 +01:00
meili-bors[bot]	a0e713c4e7	Merge #4277 4277: Update mini-dashboard to v0.2.12 r=curquiza a=mdubus # Pull Request ## Related issue Fixes #4276 ## What does this PR do? Upgrade mini-dashboard to version 0.2.12 ([see changes](https://github.com/meilisearch/mini-dashboard/releases/tag/v0.2.12)) ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Morgane Dubus <30866152+mdubus@users.noreply.github.com>	2023-12-21 11:03:46 +00:00
meili-bors[bot]	d4cb0a885b	Merge #4275 4275: Flatten settings r=dureuill a=dureuill # Pull Request ## Related issue Initial internal feedback seems to indicate that the current shape of the `embedders` setting is undesirable: it has too much depth. This PR changes this by flattening the structure of the embedders to the following: ```json5 // NEW structure "embedders": { // still starts with the embedder name "default": { "source": "huggingFace", // now a string // properties of the source are all at the same level as the source "model": "sentence-transformers/paraphrase-multilingual-MiniLM-L12-v2", "revision": "a9c555277f9bcf24f28fa5e092e665fc6f7c49cd", "documentTemplate": "A product titled '{{doc.title}}'" // now a string } } ``` By comparison, the old structure was: ```json5 // PREVIOUS version, no longer working with this PR "embedders": { // still starts with the embedder name "default": { "source": { "huggingFace": { "model": "sentence-transformers/paraphrase-multilingual-MiniLM-L12-v2", "revision": "a9c555277f9bcf24f28fa5e092e665fc6f7c49cd" }, "documentTemplate": { "template": "A product titled '{{doc.title}}'" // now a string } } } ``` The fields that are accepted in the new version of the `embedders` setting are depending on the value of the `source` field: ```json5 // huggingFace "embedders": { "default": { "source": "huggingFace", "model": "sentence-transformers/paraphrase-multilingual-MiniLM-L12-v2", "revision": "a9c555277f9bcf24f28fa5e092e665fc6f7c49cd", "documentTemplate": "A product titled '{{doc.title}}'" } } // openAi "embedders": { "default": { "source": "openAi", "model": "text-embedding-ada-002", "apiKey": "open_ai_api_key", "documentTemplate": "A product titled '{{doc.title}}'" } } // userProvided "embedders": { "default": { "source": "userProvided", "dimensions": 42, // mandatory } } ``` ## What does this PR do? - Flatten the settings structure - Validate the prompt earlier to return a synchronous error on setting change rather than in the failing task - Make it an error to pass a field for the wrong source (see above for allowed fields for each source) - Not changed: It is still an error not to pass `dimensions` to the `userProvided` embedder - If `source` was specified in the settings, validate the setting early to return a synchronous error in case of a missing mandatory field for the userProvided source (dimensions) or a forbidden field for the specified source. - If `source` was not specified in the settings, still validate the setting, but only at indexing time, by using the source stored in the DB. - Resets all values if the source changes, even if the user did not reset them explicitly. ## PR checklist Please check if your PR fulfills the following requirements: - [ ] Change the public facing guide for using the API - [ ] Change examples of use in the changelog Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-12-21 09:58:01 +00:00
Morgane Dubus	f52dee2b3b	Update Cargo.toml Update mini-dashboard with v0.2.12	2023-12-21 09:53:13 +01:00
Louis Dureuil	0bf879fb88	Fix warning on rust stable	2023-12-20 17:48:09 +01:00
Louis Dureuil	6ff81de401	Fix tests	2023-12-20 17:16:46 +01:00
Louis Dureuil	2e4c9651df	Validate settings in route	2023-12-20 17:16:46 +01:00
Louis Dureuil	ec9649c922	Add function to validate settings in Meilisearch, to be used in the routes	2023-12-20 17:16:46 +01:00
Louis Dureuil	9123370e90	Validate fused settings in settings task after fusing with existing setting	2023-12-20 17:16:46 +01:00
Louis Dureuil	14b396d302	Add new errors	2023-12-20 17:16:45 +01:00
Louis Dureuil	393216bf30	Flatten embedders settings	2023-12-20 17:16:43 +01:00
Louis Dureuil	e249e4db7b	Change Setting::apply function signature	2023-12-20 17:15:24 +01:00
meili-bors[bot]	de2ca7006e	Merge #4272 4272: Don't pass default revision when the model is explicitly set in config r=Kerollmops a=dureuill # Pull Request ## Related issue Fixes #4271 ## What does this PR do? - When the `model` is explicitly set in the `embedders` setting, we reset the `revision` to `None`, such that if the user doesn't specify a revision, the head of the model repository is chosen. - Not changed: If the user specifies a revision, it applies, like previously. - Not changed: If the user doesn't specify a model, the default model with the default revision applies, like previously. ## Manual testing on a fresh DB 1. Enable experimental feature: ```sh curl \ -X PATCH 'http://localhost:7700/experimental-features/' \ -H 'Content-Type: application/json' -H 'Authorization: Bearer foo' \ --data-binary '{ "vectorStore": true }' ``` 2. Send settings with a specified model but no specified revision: ```sh curl \ -X PATCH 'http://localhost:7700/indexes/products/settings' \ -H 'Content-Type: application/json' --data-binary \ '{ "embedders": { "default": { "source": { "huggingFace": { "model": "sentence-transformers/paraphrase-multilingual-MiniLM-L12-v2" } }, "documentTemplate": { "template": "A product titled '{{doc.title}}'"} } } }' ``` 3. Check that the task was successful: ```sh curl 'http://localhost:7700/tasks/0' {"uid":0,"indexUid":"products","status":"succeeded","type":"settingsUpdate","canceledBy":null,"details":{"embedders":{"default":{"source":{"huggingFace":{"model":"sentence-transformers/paraphrase-multilingual-MiniLM-L12-v2"}},"documentTemplate":{"template":"A product titled {{doc.title}}"}}}},"error":null,"duration":"PT0.001892S","enqueuedAt":"2023-12-20T09:17:01.73789Z","startedAt":"2023-12-20T09:17:01.73854Z","finishedAt":"2023-12-20T09:17:01.740432Z"} ``` 4. Send documents to index: ```sh curl 'https://localhost:7700/indexes/products/documents' -H 'Content-Type: application/json' --data-binary '{"id": 0, "title": "Best product"}' ``` Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-12-20 14:27:51 +00:00
Louis Dureuil	333ce12eb2	Fixed issue where the default revision is always the one we picked for the default model	2023-12-20 10:17:49 +01:00
meili-bors[bot]	fb9db1eba6	Merge #4269 4269: Remove dependency that requires libstdc++ r=dureuill a=dureuill Removes the dependency that caused the additional runtime dependency on libstdc++ by disabling the default features of the hf tokenizer. ## Discussion - This removes a feature that is using a C++ dependency and is supposed to accelerate the tokenizer. As the tokenizer is likely to be a significant bottleneck for embedding texts using a HF model, this is an issue. - We should at least rerun the movies vector indexing and check that it still works correctly and that it has a runtime in the ballpark of what it used to be. Co-authored-by: Louis Dureuil <louis.dureuil@xinra.net>	2023-12-19 12:26:48 +00:00
Clément Renault	fa2b96b9a5	Add an Authorization Header along with the webhook calls	2023-12-19 12:18:45 +01:00
Tamo	19736cefe8	add the analytics	2023-12-19 10:36:04 +01:00
Tamo	4fb25b8782	fix clippy	2023-12-19 10:35:51 +01:00
Tamo	c83a33017e	stream and chunk the data	2023-12-19 10:35:51 +01:00
Tamo	be72326c0a	gzip the tasks	2023-12-19 10:35:51 +01:00
Tamo	547379abb0	parse the url correctly	2023-12-19 10:35:51 +01:00
Tamo	0b2fff27f2	update and fix the test	2023-12-19 10:35:51 +01:00
Tamo	3adbc2b942	return a task view instead of a task	2023-12-19 10:35:51 +01:00
Tamo	fbea721378	add a first working test with actixweb	2023-12-19 10:35:51 +01:00
Tamo	391eb72137	start writing a test with actix but it doesn't works	2023-12-19 10:35:50 +01:00
Tamo	d78ad51082	Implement the webhook	2023-12-19 10:35:50 +01:00
Tamo	1956045a06	add the option	2023-12-19 10:23:56 +01:00
Louis Dureuil	b2193e612f	Revert "Add libstdc++ in Dockerfile" as it is no longer needed This reverts commit `9df8cfc013`.	2023-12-18 22:17:29 +01:00
Louis Dureuil	942d49314c	Remove dependency that requires libstdc++	2023-12-18 22:17:18 +01:00
meili-bors[bot]	9a846e82bc	Merge #4268 4268: Add libstdc++ in Dockerfile r=curquiza a=sanders41 # Pull Request ## Related issue Fixes #4267 ## What does this PR do? - Add libstdc++ in the Dockerfile ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Paul Sanders <psanders1@gmail.com>	2023-12-18 18:35:53 +00:00
Paul Sanders	9df8cfc013	Add libstdc++ in Dockerfile	2023-12-18 13:05:46 -05:00
dependabot[bot]	d868131bb7	Bump rustls-webpki from 0.101.3 to 0.101.7 Bumps [rustls-webpki](https://github.com/rustls/webpki) from 0.101.3 to 0.101.7. - [Release notes](https://github.com/rustls/webpki/releases) - [Commits](https://github.com/rustls/webpki/compare/v/0.101.3...v/0.101.7) --- updated-dependencies: - dependency-name: rustls-webpki dependency-type: indirect ... Signed-off-by: dependabot[bot] <support@github.com>	2023-12-18 14:57:38 +00:00
meili-bors[bot]	248aaa6d45	Merge #4262 4262: Update version for the next release (v1.6.0) in Cargo.toml r=curquiza a=meili-bot ⚠️ This PR is automatically generated. Check the new version is the expected one and Cargo.lock has been updated before merging. Co-authored-by: curquiza <curquiza@users.noreply.github.com>	2023-12-18 14:00:19 +00:00
curquiza	50d6317ec0	Update version for the next release (v1.6.0) in Cargo.toml	2023-12-18 13:57:46 +00:00
meili-bors[bot]	b734bd9891	Merge #4261 4261: Set rust toolchain to 1.71.1 in dockerfile r=curquiza a=dureuill Fixes docker [CI](https://github.com/meilisearch/meilisearch/actions/workflows/publish-docker-images.yml) Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-12-18 12:32:26 +00:00
Louis Dureuil	9800d5a103	Set rust toolchain to 1.71.1 in dockerfile	2023-12-18 10:59:25 +01:00
meili-bors[bot]	7c4ed07617	Merge #4257 4257: Change proximity precision settings r=dureuill a=ManyTheFish - [x] Add proximity_precision value into the analytics - [x] Change the naming of `attributeScale` and `wordScale` into `byAttribute` and `byWord` - [x] Remove proximityPrecision from the experimental feature Co-authored-by: ManyTheFish <many@meilisearch.com> Co-authored-by: Many the fish <many@meilisearch.com>	2023-12-18 09:07:28 +00:00
ManyTheFish	3a99a555a2	Fix experimental features snapshots in tests	2023-12-18 10:05:51 +01:00
Many the fish	9e1b458010	Merge branch 'main' into change-proximity-precision-settings	2023-12-18 09:08:47 +01:00
meili-bors[bot]	2aede03bc2	Merge #4226 4226: Hybrid search r=dureuill a=dureuill Allows to perform hybrid search requests that combine the results of semantic and keyword search and automatically generate embeddings. ## How to use See [feature description](https://meilisearch.notion.site/v1-6-Hybrid-Search-Embedders-ea42c82f90cc4bc0be1eeb917c1118c8) ## Changes - work is based on #4213 - milli::new search now takes an input universe directly, rather than computing it from a filter. This adds flexibility to require results on a subset of documents - vector search is now a regular ranking rule (akin to sort and geosort) and reports its score as a ScoreDetail - separate keyword search and vector search functions, vector search now respects (geo)sort ranking rules - add automatic embedding - add hybrid search Co-authored-by: Louis Dureuil <louis@meilisearch.com> Co-authored-by: ManyTheFish <many@meilisearch.com>	2023-12-14 16:24:56 +00:00
ManyTheFish	e741bc1c62	Add proximity_precision value into the analytics	2023-12-14 16:48:06 +01:00
ManyTheFish	6425996e36	Change the naming of attributeScale and wordScale into byAttribute and byWord	2023-12-14 16:31:00 +01:00
Louis Dureuil	eb5cb91da2	Switch default from hf to openai	2023-12-14 16:19:46 +01:00
Louis Dureuil	87bba98bd8	Various changes - fixed seed for arroy - check vector dimensions as soon as it is provided to search - don't embed whitespace	2023-12-14 16:08:42 +01:00
Louis Dureuil	217105b7da	hybrid search uses semantic ratio, error handling	2023-12-14 16:08:42 +01:00
ManyTheFish	1b7c164a55	Pass the semantic ratio to milli	2023-12-14 16:08:42 +01:00
ManyTheFish	f3f3944469	Fix error checking	2023-12-14 16:08:42 +01:00
ManyTheFish	93dcbf598d	Deserialize semantic ratio	2023-12-14 16:08:42 +01:00
ManyTheFish	ac68f33194	Add simple test	2023-12-14 16:08:42 +01:00
ManyTheFish	9991152bbe	Add TODOs	2023-12-14 16:08:42 +01:00
Louis Dureuil	a4536b1381	Small adjustments to respect the spec	2023-12-14 16:08:42 +01:00
Louis Dureuil	5b51cb04af	Remove some settings	2023-12-14 16:08:42 +01:00
Louis Dureuil	3c1a14f1cd	Add settings routes	2023-12-14 16:08:42 +01:00
Louis Dureuil	b8e4709dfa	Remove prompt strategy and fallback	2023-12-14 16:08:41 +01:00
Louis Dureuil	806e5b6899	Tests pass	2023-12-14 16:08:41 +01:00
Louis Dureuil	61bd2fb7a9	Update arroy	2023-12-14 16:08:41 +01:00
Louis Dureuil	e0cc775dc4	Various changes - DistributionShift in Search object (to be set from model in embed?) - Fix issue where embedder index wasn't computed at search time - Accept as default embedder either the "default" one, or the only embedder when there is only one	2023-12-14 16:08:41 +01:00
Louis Dureuil	12940d79a9	WIP - manual embedder - multi embedders OK - clippy + tests OK	2023-12-14 16:08:41 +01:00
Louis Dureuil	922a640188	WIP multi embedders fixed template bugs	2023-12-14 16:08:41 +01:00
Louis Dureuil	abbe131084	Cosmetic change	2023-12-14 16:08:41 +01:00
Louis Dureuil	d4715e0c4d	Fix same vector sort bug	2023-12-14 16:08:41 +01:00
Louis Dureuil	11e2a2c1aa	Fix geosort bug	2023-12-14 16:08:41 +01:00
Louis Dureuil	65e49b7092	Remove stuff, add distribution shift (WIP)	2023-12-14 16:08:38 +01:00
Louis Dureuil	e56f160032	Actually pass embedders on reindex	2023-12-14 16:07:49 +01:00
Louis Dureuil	687d92f217	prompt bifluor+	2023-12-14 16:07:49 +01:00
Louis Dureuil	fb539f61fe	WIP	2023-12-14 16:07:49 +01:00
Louis Dureuil	cb4ebe163e	WIP	2023-12-14 16:07:49 +01:00
Louis Dureuil	dde3a04679	WIP arroy integration	2023-12-14 16:07:49 +01:00
Louis Dureuil	13c2c6c16b	Small commit to add hybrid search and autoembedding	2023-12-14 16:07:48 +01:00
Louis Dureuil	21bcf32109	Add candle and hg_hub, updating a lot of deps in the process	2023-12-14 16:07:48 +01:00
ManyTheFish	35e1981488	Remove proximityPrecision form the experimental feature	2023-12-14 15:52:42 +01:00
meili-bors[bot]	e0f712b9d3	Merge #4254 4254: Bring back v1.5.1 changes into main r=ManyTheFish a=Kerollmops This pull request brings back changes from the _release-v1.5.1_ branch into _main_. Co-authored-by: ManyTheFish <many@meilisearch.com> Co-authored-by: meili-bors[bot] <89034592+meili-bors[bot]@users.noreply.github.com> Co-authored-by: curquiza <curquiza@users.noreply.github.com> Co-authored-by: Clément Renault <clement@meilisearch.com>	2023-12-14 09:41:57 +00:00
Clément Renault	56571f762a	Merge remote-tracking branch 'origin/main' into tmp-release-v1.5.1	2023-12-13 11:57:01 +01:00
Clément Renault	005800634d	Merge pull request #4249 from meilisearch/flag-limit-batch-size Introduce parameters to limit the number of batched tasks	2023-12-13 10:32:14 +01:00
Clément Renault	976af4fa8f	Add the default commented experimental batched tasks limit parameter to the config file	2023-12-12 10:59:00 +01:00
Clément Renault	99fec27788	Make the --max-number-of-batched-tasks argument experimental	2023-12-12 10:55:39 +01:00
meili-bors[bot]	afa8f273a8	Merge #4250 4250: Update version for the next release (v1.5.1) in Cargo.toml r=dureuill a=meili-bot ⚠️ This PR is automatically generated. Check the new version is the expected one and Cargo.lock has been updated before merging. Co-authored-by: curquiza <curquiza@users.noreply.github.com>	2023-12-12 08:26:06 +00:00
curquiza	4b644f6bc0	Update version for the next release (v1.5.1) in Cargo.toml	2023-12-11 17:15:11 +00:00
Clément Renault	7e259cb0d2	Expose the --max-number-of-batched-tasks argument	2023-12-11 16:08:39 +01:00
meili-bors[bot]	0fbc1511d7	Merge #4225 4225: [EXP] Let the user customize the proximity precision r=dureuill a=ManyTheFish # Pull Request This PR introduces a new setting `proximityPrecision` allowing the user to trade indexing time with search precision on proximity-based features: - proximity ranking rules - multi-word synonyms - phrase search - split-words I put the API PRD below: https://www.notion.so/meilisearch/3988b345b5b248948a4a0dc5932a18ce?v=45d79150adb84b0aa27826ff6da2e029&p=aa69c2bab2c3402bab9340ae4def4577&pm=s ## Related issue Fixes #4187 Co-authored-by: ManyTheFish <many@meilisearch.com>	2023-12-06 17:21:43 +00:00
ManyTheFish	c9860c7913	Small test fixes	2023-12-06 15:49:05 +01:00
ManyTheFish	03ffabe889	Add a new dump test	2023-12-06 15:49:05 +01:00
ManyTheFish	1f4fc9c229	Make the feature experimental	2023-12-06 15:49:05 +01:00
ManyTheFish	8cc3c54117	Add proximityPrecision setting in settings route	2023-12-06 15:49:05 +01:00
ManyTheFish	467b49153d	Implement proximityPrecision setting on milli side	2023-12-06 15:49:02 +01:00
ManyTheFish	0c3fa8cbc4	Add tests on proximityPrecision setting	2023-12-06 14:59:23 +01:00
ManyTheFish	bddc168d83	List TODOs	2023-12-06 14:59:23 +01:00
meili-bors[bot]	84a36002d7	Merge #4239 4239: Remove the actix-web dependency from milli r=dureuill a=Kerollmops Just remove actix-web from milli. Co-authored-by: Clément Renault <clement@meilisearch.com>	2023-11-29 10:19:40 +00:00
meili-bors[bot]	c95d68e244	Merge #4233 4233: Add test reproducing #4232 r=dureuill a=ManyTheFish - add a test reproducing the bug - fix the bug by creating 2 different restricting lists of attributes, one for the exact attributes, and the other for the tolerant attributes ## Related issue Fixes #4232 Co-authored-by: ManyTheFish <many@meilisearch.com>	2023-11-29 08:47:17 +00:00
ManyTheFish	3b3fa38f27	Put the restrict list in a sub-struct	2023-11-28 18:37:57 +01:00
Clément Renault	170e063b80	Remove the actix-web dependency from milli	2023-11-28 17:19:57 +01:00
ManyTheFish	d6c2ee15a9	Filter on attributes before computing the docids when attribute restriction is on	2023-11-28 14:55:29 +01:00
meili-bors[bot]	6376c342c1	Merge #4223 4223: Update to heed 0.20 r=dureuill a=Kerollmops This PR brings the v0.20-alpha.9 version of heed into Meilisearch 🎉 The main goal is to test it in a real environment to make the necessary changes if needed. We also want to merge it as soon as possible during the pre-release phase to ensure we catch bugs before the release. Most of the calls to heed are the same as before, except: - The `PolyDatabase` has been replaced with a `Database<Unspecified, Unspecified>`. We replaced the `get<T, U>()` by a `remap<T, U>().get()` calls. - The `Database` `append(...)` method has been replaced with a `put_with_flags(PutFlags::APPEND, ...)`. - The `RwTxn<'e, 'p>` has been simplified into a `RwTxn<'e>`. - The `BytesEncode/Decode` traits return a `Result<_, BoxedError>` instead of an `Option<_>`. - We no longer need to wrap and unwrap the `BEU32` integer when storing/getting them from heed. ### TODO - [x] Create actual, simple error types instead of using strings in the codecs. ### Follow-up work - Move the codecs into another member crate (we depend on the uuid one in the meilitool crate). - Display the internal decoding error in the `SerializationError` internal error variant. Co-authored-by: Clément Renault <clement@meilisearch.com>	2023-11-28 13:39:44 +00:00
Clément Renault	5b563f872b	Move the clippy attribute on the problematic part of the code	2023-11-28 14:37:58 +01:00
Clément Renault	ec9b52d608	Rename copy_to_path to copy_to_file	2023-11-28 14:32:30 +01:00
Clément Renault	34c67ac389	Remove the possibility to fail fetching the env info	2023-11-28 14:31:23 +01:00
Clément Renault	d050c9b4ae	Only remap the main database once	2023-11-28 14:27:30 +01:00
Clément Renault	7dd1226faf	Clarify an unreachable unwrap	2023-11-28 14:26:31 +01:00
Clément Renault	1575456594	Further reduce an async block	2023-11-28 14:23:32 +01:00
Clément Renault	add2ceef67	Introduce error types to avoid panics	2023-11-28 14:21:49 +01:00
Clément Renault	548c8247c2	Create and use real error types in the codecs	2023-11-28 10:11:17 +01:00
meili-bors[bot]	181ca48482	Merge #4234 4234: Fix puffin in the index scheduler r=dureuill a=irevoire Currently, we can't compile the index scheduler without this feature. It could be cool to specify the dependencies in the main workspace cargo toml like quickwit does to avoid this kind of error in the future; https://github.com/quickwit-oss/quickwit/blob/main/quickwit/Cargo.toml#L41 Co-authored-by: Tamo <tamo@meilisearch.com>	2023-11-28 08:23:48 +00:00
Tamo	5751f5c640	fix puffin in the index scheduler	2023-11-27 15:18:33 +01:00
Clément Renault	d32eb11329	Move to the v0.20.0-alpha.9 of heed	2023-11-27 11:52:22 +01:00
ManyTheFish	dc07790133	Add test reproducing #4232	2023-11-27 11:39:11 +01:00
meili-bors[bot]	3d23b388bc	Merge #4231 4231: Fixed payload limit setting being ignored for delete documents by batch r=Kerollmops a=Karribalu # Pull Request ## Related issue Fixes #4224 ## What does this PR do? - Added http_payload_size_limit to JsonConfig to allow deleting documents in batches with a payload size greater than 2MB, which is the default limit set in the JsonConfig crate. ## PR checklist Please check if your PR fulfills the following requirements: - [Y] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [Y] Have you read the contributing guidelines? - [Y] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: karribalu <karri.balu123456@gmail.com>	2023-11-27 09:26:21 +00:00
karribalu	85626cff8e	Fixed payload limit setting being ignored for delete documents by batch route	2023-11-25 18:41:16 +00:00
Clément Renault	58dac8af42	Remove the panics and unwraps	2023-11-23 15:00:48 +01:00
Clément Renault	0dbf1a16ff	Make clippy happy	2023-11-23 14:11:38 +01:00
Clément Renault	462b4c0080	Fix the tests	2023-11-23 12:07:35 +01:00
Clément Renault	0d4482625a	Make the changes to use heed v0.20-alpha.6	2023-11-23 11:43:58 +01:00
Clément Renault	56a0d91ecd	Update the heed dependency and lock file	2023-11-22 15:11:09 +01:00
meili-bors[bot]	b366acdae6	Merge #4220 4220: Bring back changes from v1.5.0 into main r=dureuill a=Kerollmops This will bring the fixes from v1.5.0 into main. By [following this guide](https://github.com/meilisearch/engine-team/blob/main/resources/meilisearch-release.md#after-the-release) I decided to create a temporary branch to fix the git conflicts and merge into main afterward. Co-authored-by: curquiza <curquiza@users.noreply.github.com> Co-authored-by: Vivek Kumar <vivek.26@outlook.com> Co-authored-by: Louis Dureuil <louis.dureuil@gmail.com> Co-authored-by: meili-bors[bot] <89034592+meili-bors[bot]@users.noreply.github.com> Co-authored-by: ManyTheFish <many@meilisearch.com> Co-authored-by: Tamo <tamo@meilisearch.com> Co-authored-by: Clément Renault <clement@meilisearch.com> Co-authored-by: Louis Dureuil <louis.dureuil@xinra.net> Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-11-22 07:46:22 +00:00
Clément Renault	7cb7e37ba8	Merge branch 'main' into tmp-release-v1.5.0	2023-11-21 16:30:46 +01:00
meili-bors[bot]	33b7c574ea	Merge #4090 4090: Diff indexing r=ManyTheFish a=ManyTheFish This pull request aims to reduce the indexing time by computing a difference between the data added to the index and the data removed from the index before writing in LMDB. ## Why focus on reducing the writings in LMDB? The indexing in Meilisearch is split into 3 main phases: 1) The computing or the extraction of the data (Multi-threaded) 2) The writing of the data in LMDB (Mono-threaded) 3) The processing of the prefix databases (Mono-threaded) see below: ![Capture d’écran 2023-09-28 à 20 01 45](https://github.com/meilisearch/meilisearch/assets/6482087/51513162-7c39-4244-978b-2c6b60c43a56) Because the writing is mono-threaded, it represents a bottleneck in the indexing, reducing the number of writes in LMDB will reduce the pressure on the main thread and should reduce the global time spent on the indexing. ## Give Feedback We created [a dedicated discussion](https://github.com/meilisearch/meilisearch/discussions/4196) for users to try this new feature and to give feedback on bugs or performance issues. ## Technical approach ### Part 1: merge the addition and the deletion process This part: a) Aims to reduce the time spent on indexing only the filterable/sortable fields of documents, for example: - Updating the number of "likes" or "stars" of a song or a movie - Updating the "stock count" or the "price" of a product b) Aims to reduce the time spent on writing in LMDB which should reduce the global indexing time for the highly multi-threaded machines by reducing the writing bottleneck. c) Aims to reduce the average time spent to delete documents without having to keep the soft-deleted documents implementation - [x] Create a preprocessing function that creates the diff-based documents chuck (`OBKV<fid, OBKV<AddDel, value>>`) - [x] and clearly separate the faceted fields and the searchable fields in two different chunks - Change the parameters of the input extractor by taking an `OBKV<fid, OBKV<AddDel, value>>` instead of `OBKV<fid, value>`. - [x] extract_docid_word_positions - [x] extract_geo_points - [x] extract_vector_points - [x] extract_fid_docid_facet_values - Adapt the searchable extractors to the new diff-chucks - [x] extract_fid_word_count_docids - [x] extract_word_pair_proximity_docids - [x] extract_word_position_docids - [x] extract_word_docids - Adapt the facet extractors to the new diff-chucks - [x] extract_facet_number_docids - [x] extract_facet_string_docids - [x] extract_fid_docid_facet_values - [x] FacetsUpdate - [x] Adapt the prefix database extractors ⚠️ ⚠️ - [x] Make the LMDB writer remove the document_ids to delete at the same time the new document_ids are added - [x] Remove document deletion pipeline - [x] remove `new_documents_ids` entirely and `replaced_documents_ids` - [x] reuse extracted external id from transform instead of re-extracting in `TypedChunks::Documents` - [x] Remove deletion pipeline after autobatcher - [x] remove autobatcher deletion pipeline - [x] everything uses `IndexOperation::DocumentOperation` - [x] repair deletion by internal id for filter by delete - [x] Improve the deletion via internal ids by avoiding iterating over the whole set of external document ids. - [x] Remove soft-deleted documents #### FIXME - [x] field distribution is not correctly updated after deletion - [x] missing documents in the tests of tokenizer_customization ### Part 2: Only compute the documents field by field This part aims to reduce the global indexing time for any kind of partial document modification on any size of machine from the mono-threaded one to the highly multi-threaded one. - [ ] Make the preprocessing function only send the fields that changed to the extractors - [ ] remove the `word_docids` and `exact_word_docids` database and adapt the search (⚠️ could impact the search performances) - [ ] replace the `word_pair_proximity_docids` database with a `word_pair_proximity_fid_docids` database and adapt the search (⚠️ could impact the search performances) - [ ] Adapt the prefix database extractors ⚠️ ⚠️ ## Technical Concerns - The part 1 implementation could increase the indexing time for the smallest machines (with few threads) by increasing the extracting time (multi-threaded) more than the writing time (mono-threaded) - The part 2 implementation needs to change the databases which could have a significant impact on the search performances - The prefix databases are a bit special to process and may be a pain to adapt to the difference-based indexing Co-authored-by: ManyTheFish <many@meilisearch.com> Co-authored-by: Clément Renault <clement@meilisearch.com> Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-11-21 09:44:38 +00:00
ManyTheFish	d3575fb028	Make into_del_add_obkv parameters more human readable	2023-11-20 16:10:39 +01:00
ManyTheFish	39cbb499c2	Small fixes	2023-11-20 10:20:39 +01:00
ManyTheFish	ebef6bc24d	Simplify documents database writing	2023-11-20 10:14:57 +01:00
ManyTheFish	d59b7db8d0	remove unused code	2023-11-20 10:10:45 +01:00
ManyTheFish	263e825619	Fix typos in comments	2023-11-20 10:06:29 +01:00
Clément Renault	69354a6144	Add the benchmarck name to the bot message	2023-11-15 13:56:54 +01:00
Many the fish	b0adc73ce6	Merge pull request #4207 from meilisearch/diff-indexing-prefix-databases Diff indexing prefix databases	2023-11-14 16:04:05 +01:00
meili-bors[bot]	2b5d9042d1	Merge #4208 4208: Makes the dump cancellable r=Kerollmops a=irevoire # Pull Request Make the dump tasks cancellable even when they have already started processing. ## Related issue Fixes https://github.com/meilisearch/meilisearch/issues/4157 Co-authored-by: Tamo <tamo@meilisearch.com>	2023-11-14 13:31:45 +00:00
Tamo	5b57fbab08	makes the dump cancellable	2023-11-14 11:23:13 +01:00
meili-bors[bot]	72d3fa4898	Merge #4203 4203: Extract external document docids from docs on deletion by filter r=Kerollmops a=dureuill This fixes some of the performance regression observed on `diff-indexing` when doing delete-by-filter with a filter matching many documents. To delete 19 768 771 documents (hackernews dataset, all documents matching `type = comment`), here are the observed time: \|branch (commit sha1sum)\|time\|speed-down factor (lower is better)\| \|--\|--\|--\| \|`main` (`48865470d7`)\|1212.885536s (~20min)\|x1.0 (baseline)\| \|`diff-indexing` (`523519fdbf`)\|5385.550543s (90min)\|x4.44\| \|`diff-indexing-extract-primary-key`(`f8289cd974`)\|2582.323324s (43min) \| x2.13\| So we're still suffering a speed-down of x2.13, but that's much better than x4.44. --- Changes: - Refactor the logic of PrimaryKey extraction to a struct - Add a trait to abstract the extraction of field id from a name between `DocumentBatch` and `FieldIdMap`. - Add `Index::external_id_of` to get the external ids of a bitmap of internal ids. - Use this new method to add new Transform and Batch methods to remove documents that are known to be from the DB. - Modify delete-by-filter to use the new method Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-11-13 13:02:10 +00:00
Louis Dureuil	772964125d	Factor removal of document from DB	2023-11-13 13:51:22 +01:00
Louis Dureuil	378deb0bef	Rename trait	2023-11-13 13:38:36 +01:00
ManyTheFish	1f36410541	Update tests	2023-11-13 13:36:39 +01:00
meili-bors[bot]	b11f85a635	Merge #4205 4205: Prevent search hang on the processing index r=Kerollmops a=dureuill Fixes #4206, an issue originally [reported on Discord](https://discord.com/channels/1006923006964154428/1148983671026618579/1148983671026618579) where having parallel search requests on more indexes than the index cache capacity would cause search requests on the currently updating index to hang until the index is done updating. ## Test setup - Create 20 empty indexes by sending settings to them - repeatedly send placeholder search requests to each of the indexes in a loop - Create another index and send a significant batch of documents to index. - Attempt to perform a search request on that last index. - Before this PR, the search request hangs while the index update task is processing - After this PR, the search request respond immediately even while the index update task is processing ## Changes - When getting the handle to an index for some potentially long running batches of tasks, save it in the index scheduler. - Drop the handle from the index-scheduler when the task is done so that we don't leak indexes. - When getting an index from outside the task queue processor, check if there is such an handle matching the requested index. If so, skip the cache entirely and clone the handle. Co-authored-by: Louis Dureuil <louis.dureuil@xinra.net> Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-11-13 10:36:01 +00:00
Louis Dureuil	a2d6dc8571	Fix typo, remove caching for the change of index	2023-11-13 10:44:36 +01:00
meili-bors[bot]	ee1701157f	Merge #4204 4204: Throw error when the vector search is sent with the wrong size r=Kerollmops a=dureuill # Pull Request ## Related issue Fixes #4201 Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-11-13 09:43:20 +00:00
Louis Dureuil	8c649d8061	Throw error when the vector search is sent with the wrong size	2023-11-13 09:57:42 +01:00
Louis Dureuil	492fc086f0	cargo fmt	2023-11-12 21:53:11 +01:00
Louis Dureuil	a2d0c73b41	Save the currently updating index so that the search can access it at all times	2023-11-10 10:52:03 +01:00
Louis Dureuil	264b10ec20	Fixup documentation	2023-11-09 16:23:20 +01:00
Louis Dureuil	825257da76	Use more efficient method for deletion in benchmarks	2023-11-09 16:13:15 +01:00
Louis Dureuil	f8289cd974	Use it from delete-by-filter	2023-11-09 14:23:15 +01:00
Louis Dureuil	3053e01c05	Batch::remove_documents_from_db_no_batch	2023-11-09 14:23:02 +01:00
Louis Dureuil	b11c2afac0	Index::external_id_of	2023-11-09 14:22:43 +01:00
Louis Dureuil	9cef800b2a	Enrich uses the new type	2023-11-09 14:22:05 +01:00
Louis Dureuil	db2fb86b8b	Extract PrimaryKey logic to a type	2023-11-09 14:19:16 +01:00
ManyTheFish	882ab9cc85	remove warnings	2023-11-09 11:35:33 +01:00
ManyTheFish	5a9c96e1db	Compute word integer prefix cache	2023-11-09 11:34:26 +01:00
ManyTheFish	70ce40828c	Compute word docids prefix cache	2023-11-08 17:01:00 +01:00
ManyTheFish	688266c83e	Remove word pair proximity prefix cache and compute it at search time	2023-11-08 14:16:01 +01:00
ManyTheFish	6dab826908	Reactivate prefix databases	2023-11-08 13:58:01 +01:00
ManyTheFish	1e2fbc6a42	revert "REVERT ME: ignore prefix pair databases tests" This reverts commit `1b2ea6cf19`.	2023-11-08 11:50:52 +01:00
Many the fish	523519fdbf	Merge pull request #4195 from meilisearch/diff-indexing-remove-from-batch Remove `IndexOperation::DocumentDeletion`	2023-11-08 10:29:49 +01:00
Louis Dureuil	ef6fa10f7a	Remove `IndexOperation::DocumentDeletion`	2023-11-06 12:16:15 +01:00
Louis Dureuil	620fee35f9	Fix benches	2023-11-06 11:56:46 +01:00
Louis Dureuil	cbaa54cafd	Fix clippy issues	2023-11-06 11:19:31 +01:00
Louis Dureuil	1bccf2079e	Correctly mark non-tests as non-tests	2023-11-06 11:03:56 +01:00
ManyTheFish	1b2ea6cf19	REVERT ME: ignore prefix pair databases tests	2023-11-06 10:46:22 +01:00
Louis Dureuil	1ad1fcc8c8	Remove all warnings	2023-11-06 10:31:14 +01:00
meili-bors[bot]	48865470d7	Merge #4191 4191: Remove banner r=Kerollmops a=curquiza Co-authored-by: Clémentine U. - curqui <clementine@meilisearch.com>	2023-11-02 17:14:23 +00:00
Clémentine U. - curqui	c810df4d9f	Update README.md	2023-11-02 17:40:18 +01:00
ManyTheFish	87610a5f98	Don't try to delete a document that is not in the database	2023-11-02 16:49:03 +01:00
Many the fish	2544bc1416	Merge pull request #4160 from meilisearch/diff-indexing-vector-points Diff Indexing for the vector points	2023-11-02 16:01:51 +01:00
Clément Renault	ff522c919d	Fix the vector extractions for the diff indexing	2023-11-02 15:58:08 +01:00
Many the fish	1c39459cf4	Merge pull request #4179 from meilisearch/diff-indexing-fix-nested-primary-key Diff indexing fix nested primary key	2023-11-02 15:39:50 +01:00
ManyTheFish	bf0651f23c	Implement iter method on ExternalDocumentsIds	2023-11-02 15:38:00 +01:00
ManyTheFish	5b20e625f3	fix merge	2023-11-02 15:31:37 +01:00
ManyTheFish	bc51d6157a	Fix transform reindexing path	2023-11-02 15:26:20 +01:00
ManyTheFish	1b4ff991c0	update typed chunks	2023-11-02 15:26:20 +01:00
ManyTheFish	4b64c33aa2	update vector extractor	2023-11-02 15:26:20 +01:00
ManyTheFish	12323d610e	Change the original document sorter key from the internal docid to a concatenation of the internal and the external docid	2023-11-02 15:26:20 +01:00
Clément Renault	44e9033b3a	Merge pull request #4181 from meilisearch/diff-indexing-parallel-transform Use rayon to sort entries in parallel	2023-11-02 15:16:10 +01:00
Clément Renault	4d864f0702	Always sort internal Sorter entries in parallel	2023-11-02 14:47:43 +01:00
meili-bors[bot]	5e3df76699	Merge #4183 4183: Bump docker/login-action from 2 to 3 r=curquiza a=dependabot[bot] Bumps [docker/login-action](https://github.com/docker/login-action) from 2 to 3. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/docker/login-action/releases">docker/login-action's releases</a>.</em></p> <blockquote> <h2>v3.0.0</h2> <ul> <li>Node 20 as default runtime (requires <a href="https://github.com/actions/runner/releases/tag/v2.308.0">Actions Runner v2.308.0</a> or later) by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` in <a href="https://redirect.github.com/docker/login-action/pull/593">docker/login-action#593</a></li> <li>Bump <code>`@actions/core</code>` from 1.10.0 to 1.10.1 in <a href="https://redirect.github.com/docker/login-action/pull/598">docker/login-action#598</a></li> <li>Bump <code>`@aws-sdk/client-ecr</code>` and <code>`@aws-sdk/client-ecr-public</code>` to 3.410.0 in <a href="https://redirect.github.com/docker/login-action/pull/555">docker/login-action#555</a> <a href="https://redirect.github.com/docker/login-action/pull/560">docker/login-action#560</a> <a href="https://redirect.github.com/docker/login-action/pull/582">docker/login-action#582</a> <a href="https://redirect.github.com/docker/login-action/pull/599">docker/login-action#599</a></li> <li>Bump semver from 6.3.0 to 6.3.1 in <a href="https://redirect.github.com/docker/login-action/pull/556">docker/login-action#556</a></li> <li>Bump https-proxy-agent to 7.0.2 <a href="https://redirect.github.com/docker/login-action/pull/561">docker/login-action#561</a> <a href="https://redirect.github.com/docker/login-action/pull/588">docker/login-action#588</a></li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/login-action/compare/v2.2.0...v3.0.0">https://github.com/docker/login-action/compare/v2.2.0...v3.0.0</a></p> <h2>v2.2.0</h2> <ul> <li>Switch to actions-toolkit implementation by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` in <a href="https://redirect.github.com/docker/login-action/pull/409">docker/login-action#409</a> <a href="https://redirect.github.com/docker/login-action/pull/470">docker/login-action#470</a> <a href="https://redirect.github.com/docker/login-action/pull/476">docker/login-action#476</a></li> <li>Bump <code>`@aws-sdk/client-ecr</code>` and <code>`@aws-sdk/client-ecr-public</code>` to 3.347.1 in <a href="https://redirect.github.com/docker/login-action/pull/524">docker/login-action#524</a> <a href="https://redirect.github.com/docker/login-action/pull/364">docker/login-action#364</a> <a href="https://redirect.github.com/docker/login-action/pull/363">docker/login-action#363</a></li> <li>Bump minimatch from 3.0.4 to 3.1.2 in <a href="https://redirect.github.com/docker/login-action/pull/354">docker/login-action#354</a></li> <li>Bump json5 from 2.2.0 to 2.2.3 in <a href="https://redirect.github.com/docker/login-action/pull/378">docker/login-action#378</a></li> <li>Bump http-proxy-agent from 5.0.0 to 7.0.0 in <a href="https://redirect.github.com/docker/login-action/pull/509">docker/login-action#509</a></li> <li>Bump https-proxy-agent from 5.0.1 to 7.0.0 in <a href="https://redirect.github.com/docker/login-action/pull/508">docker/login-action#508</a></li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/login-action/compare/v2.1.0...v2.2.0">https://github.com/docker/login-action/compare/v2.1.0...v2.2.0</a></p> <h2>v2.1.0</h2> <ul> <li>Ensure AWS temp credentials are redacted in workflow logs by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://redirect.github.com/docker/login-action/issues/275">#275</a>)</li> <li>Bump <code>`@actions/core</code>` from 1.6.0 to 1.10.0 (<a href="https://redirect.github.com/docker/login-action/issues/252">#252</a> <a href="https://redirect.github.com/docker/login-action/issues/292">#292</a>)</li> <li>Bump <code>`@aws-sdk/client-ecr</code>` from 3.53.0 to 3.186.0 (<a href="https://redirect.github.com/docker/login-action/issues/298">#298</a>)</li> <li>Bump <code>`@aws-sdk/client-ecr-public</code>` from 3.53.0 to 3.186.0 (<a href="https://redirect.github.com/docker/login-action/issues/299">#299</a>)</li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/login-action/compare/v2.0.0...v2.1.0">https://github.com/docker/login-action/compare/v2.0.0...v2.1.0</a></p> </blockquote> </details> <details> <summary>Commits</summary> <ul> <li><a href="`343f7c4344`"><code>343f7c4</code></a> Merge pull request <a href="https://redirect.github.com/docker/login-action/issues/599">#599</a> from docker/dependabot/npm_and_yarn/aws-sdk-dependenc...</li> <li><a href="`aad0f974f2`"><code>aad0f97</code></a> chore: update generated content</li> <li><a href="`2e0cd39144`"><code>2e0cd39</code></a> build(deps): bump the aws-sdk-dependencies group with 2 updates</li> <li><a href="`203bc9c4ef`"><code>203bc9c</code></a> Merge pull request <a href="https://redirect.github.com/docker/login-action/issues/588">#588</a> from docker/dependabot/npm_and_yarn/proxy-agent-depen...</li> <li><a href="`2199648fc8`"><code>2199648</code></a> chore: update generated content</li> <li><a href="`b489376173`"><code>b489376</code></a> build(deps): bump the proxy-agent-dependencies group with 1 update</li> <li><a href="`7c309e74e6`"><code>7c309e7</code></a> Merge pull request <a href="https://redirect.github.com/docker/login-action/issues/598">#598</a> from docker/dependabot/npm_and_yarn/actions/core-1.10.1</li> <li><a href="`0ccf222961`"><code>0ccf222</code></a> chore: update generated content</li> <li><a href="`56d703e106`"><code>56d703e</code></a> Merge pull request <a href="https://redirect.github.com/docker/login-action/issues/597">#597</a> from docker/dependabot/github_actions/aws-actions/con...</li> <li><a href="`24d3b3519e`"><code>24d3b35</code></a> build(deps): bump <code>`@actions/core</code>` from 1.10.0 to 1.10.1</li> <li>Additional commits viewable in <a href="https://github.com/docker/login-action/compare/v2...v3">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=docker/login-action&package-manager=github_actions&previous-version=2&new-version=3)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) You can trigger a rebase of this PR by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` show <dependency name> ignore conditions` will show all of the ignore conditions of the specified dependency - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) </details> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2023-11-02 13:18:13 +00:00
meili-bors[bot]	02765fb267	Merge #4184 4184: Bump actions/setup-node from 3 to 4 r=curquiza a=dependabot[bot] Bumps [actions/setup-node](https://github.com/actions/setup-node) from 3 to 4. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/actions/setup-node/releases">actions/setup-node's releases</a>.</em></p> <blockquote> <h2>v4.0.0</h2> <h2>What's Changed</h2> <p>In scope of this release we changed version of node runtime for action from node16 to node20 and updated dependencies in <a href="https://redirect.github.com/actions/setup-node/pull/866">actions/setup-node#866</a></p> <p>Besides, release contains such changes as:</p> <ul> <li>Upgrade actions/checkout to v4 by <a href="https://github.com/gmembre-zenika"><code>`@gmembre-zenika</code></a>` in <a href="https://redirect.github.com/actions/setup-node/pull/868">actions/setup-node#868</a></li> <li>Update actions/checkout for documentation and yaml by <a href="https://github.com/dmitry-shibanov"><code>`@dmitry-shibanov</code></a>` in <a href="https://redirect.github.com/actions/setup-node/pull/876">actions/setup-node#876</a></li> </ul> <h2>New Contributors</h2> <ul> <li><a href="https://github.com/gmembre-zenika"><code>`@gmembre-zenika</code></a>` made their first contribution in <a href="https://redirect.github.com/actions/setup-node/pull/868">actions/setup-node#868</a></li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/actions/setup-node/compare/v3...v4.0.0">https://github.com/actions/setup-node/compare/v3...v4.0.0</a></p> <h2>v3.8.2</h2> <h2>What's Changed</h2> <ul> <li>Update semver by <a href="https://github.com/dmitry-shibanov"><code>`@dmitry-shibanov</code></a>` in <a href="https://redirect.github.com/actions/setup-node/pull/861">actions/setup-node#861</a></li> <li>Update temp directory creation by <a href="https://github.com/nikolai-laevskii"><code>`@nikolai-laevskii</code></a>` in <a href="https://redirect.github.com/actions/setup-node/pull/859">actions/setup-node#859</a></li> <li>Bump <code>`@babel/traverse</code>` from 7.15.4 to 7.23.2 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/actions/setup-node/pull/870">actions/setup-node#870</a></li> <li>Add notice about binaries not being updated yet by <a href="https://github.com/nikolai-laevskii"><code>`@nikolai-laevskii</code></a>` in <a href="https://redirect.github.com/actions/setup-node/pull/872">actions/setup-node#872</a></li> <li>Update toolkit cache and core by <a href="https://github.com/dmitry-shibanov"><code>`@dmitry-shibanov</code></a>` and <a href="https://github.com/seongwon-privatenote"><code>`@seongwon-privatenote</code></a>` in <a href="https://redirect.github.com/actions/setup-node/pull/875">actions/setup-node#875</a></li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/actions/setup-node/compare/v3...v3.8.2">https://github.com/actions/setup-node/compare/v3...v3.8.2</a></p> <h2>v3.8.1</h2> <h2>What's Changed</h2> <p>In scope of this release, the filter was removed within the cache-save step by <a href="https://github.com/dmitry-shibanov"><code>`@dmitry-shibanov</code></a>` in <a href="https://redirect.github.com/actions/setup-node/pull/831">actions/setup-node#831</a>. It is filtered and checked in the toolkit/cache library.</p> <p><strong>Full Changelog</strong>: <a href="https://github.com/actions/setup-node/compare/v3...v3.8.1">https://github.com/actions/setup-node/compare/v3...v3.8.1</a></p> <h2>v3.8.0</h2> <h2>What's Changed</h2> <h3>Bug fixes:</h3> <ul> <li>Add check for existing paths by <a href="https://github.com/dmitry-shibanov"><code>`@dmitry-shibanov</code></a>` in <a href="https://redirect.github.com/actions/setup-node/pull/803">actions/setup-node#803</a></li> <li>Resolve SymbolicLink by <a href="https://github.com/dmitry-shibanov"><code>`@dmitry-shibanov</code></a>` in <a href="https://redirect.github.com/actions/setup-node/pull/809">actions/setup-node#809</a></li> <li>Change passing logic for cache input by <a href="https://github.com/dmitry-shibanov"><code>`@dmitry-shibanov</code></a>` in <a href="https://redirect.github.com/actions/setup-node/pull/816">actions/setup-node#816</a></li> <li>Fix armv7 cache issue by <a href="https://github.com/louislam"><code>`@louislam</code></a>` in <a href="https://redirect.github.com/actions/setup-node/pull/794">actions/setup-node#794</a></li> <li>Update check-dist workflow name by <a href="https://github.com/sinchang"><code>`@sinchang</code></a>` in <a href="https://redirect.github.com/actions/setup-node/pull/710">actions/setup-node#710</a></li> </ul> <h3>Feature implementations:</h3> <ul> <li>feat: handling the case where "node" is used for tool-versions file. by <a href="https://github.com/xytis"><code>`@xytis</code></a>` in <a href="https://redirect.github.com/actions/setup-node/pull/812">actions/setup-node#812</a></li> </ul> <h3>Documentation changes:</h3> <ul> <li>Refer to semver package name in README.md by <a href="https://github.com/olleolleolle"><code>`@olleolleolle</code></a>` in <a href="https://redirect.github.com/actions/setup-node/pull/808">actions/setup-node#808</a></li> </ul> <h3>Update dependencies:</h3> <ul> <li>Update toolkit cache to fix zstd by <a href="https://github.com/dmitry-shibanov"><code>`@dmitry-shibanov</code></a>` in <a href="https://redirect.github.com/actions/setup-node/pull/804">actions/setup-node#804</a></li> <li>Bump tough-cookie and <code>`@azure/ms-rest-js</code>` by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/actions/setup-node/pull/802">actions/setup-node#802</a></li> <li>Bump semver from 6.1.2 to 6.3.1 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/actions/setup-node/pull/807">actions/setup-node#807</a></li> </ul> <!-- raw HTML omitted --> </blockquote> <p>... (truncated)</p> </details> <details> <summary>Commits</summary> <ul> <li><a href="`8f152de45c`"><code>8f152de</code></a> Update actions/checkout for documentation and yaml (<a href="https://redirect.github.com/actions/setup-node/issues/876">#876</a>)</li> <li><a href="`23755b521f`"><code>23755b5</code></a> upgrade actions/checkout to v4 (<a href="https://redirect.github.com/actions/setup-node/issues/868">#868</a>)</li> <li><a href="`54534a2a9b`"><code>54534a2</code></a> Change node version for action to node20 (<a href="https://redirect.github.com/actions/setup-node/issues/866">#866</a>)</li> <li>See full diff in <a href="https://github.com/actions/setup-node/compare/v3...v4">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=actions/setup-node&package-manager=github_actions&previous-version=3&new-version=4)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) You can trigger a rebase of this PR by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` show <dependency name> ignore conditions` will show all of the ignore conditions of the specified dependency - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) </details> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2023-11-02 11:28:03 +00:00
meili-bors[bot]	841165d529	Merge #4185 4185: Bump Swatinem/rust-cache from 2.6.2 to 2.7.1 r=curquiza a=dependabot[bot] Bumps [Swatinem/rust-cache](https://github.com/swatinem/rust-cache) from 2.6.2 to 2.7.1. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/swatinem/rust-cache/releases">Swatinem/rust-cache's releases</a>.</em></p> <blockquote> <h2>v2.7.0</h2> <h2>What's Changed</h2> <ul> <li>Fix save-if documentation in readme by <a href="https://github.com/rukai"><code>`@rukai</code></a>` in <a href="https://redirect.github.com/Swatinem/rust-cache/pull/166">Swatinem/rust-cache#166</a></li> <li>Support for <code>trybuild</code> and similar macro testing tools by <a href="https://github.com/neysofu"><code>`@neysofu</code></a>` in <a href="https://redirect.github.com/Swatinem/rust-cache/pull/168">Swatinem/rust-cache#168</a></li> </ul> <h2>New Contributors</h2> <ul> <li><a href="https://github.com/rukai"><code>`@rukai</code></a>` made their first contribution in <a href="https://redirect.github.com/Swatinem/rust-cache/pull/166">Swatinem/rust-cache#166</a></li> <li><a href="https://github.com/neysofu"><code>`@neysofu</code></a>` made their first contribution in <a href="https://redirect.github.com/Swatinem/rust-cache/pull/168">Swatinem/rust-cache#168</a></li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/Swatinem/rust-cache/compare/v2.6.2...v2.7.0">https://github.com/Swatinem/rust-cache/compare/v2.6.2...v2.7.0</a></p> </blockquote> </details> <details> <summary>Changelog</summary> <p><em>Sourced from <a href="https://github.com/Swatinem/rust-cache/blob/master/CHANGELOG.md">Swatinem/rust-cache's changelog</a>.</em></p> <blockquote> <h2>2.7.1</h2> <ul> <li>Update toml parser to fix parsing errors.</li> </ul> <h2>2.7.0</h2> <ul> <li>Properly cache <code>trybuild</code> tests.</li> </ul> </blockquote> </details> <details> <summary>Commits</summary> <ul> <li><a href="`3cf7f8cc28`"><code>3cf7f8c</code></a> 2.7.1</li> <li><a href="`e03705e031`"><code>e03705e</code></a> changelog</li> <li><a href="`b86d1c6caa`"><code>b86d1c6</code></a> bump all the other dependencies too</li> <li><a href="`f27990c89a`"><code>f27990c</code></a> Update Dependencies (<a href="https://redirect.github.com/swatinem/rust-cache/issues/172">#172</a>)</li> <li><a href="`a95ba19544`"><code>a95ba19</code></a> 2.7.0</li> <li><a href="`82c8487d00`"><code>82c8487</code></a> changelog</li> <li><a href="`67c46e7159`"><code>67c46e7</code></a> Support for <code>trybuild</code> and similar macro testing tools (<a href="https://redirect.github.com/swatinem/rust-cache/issues/168">#168</a>)</li> <li><a href="`44b6087283`"><code>44b6087</code></a> Fix save-if documentation in readme (<a href="https://redirect.github.com/swatinem/rust-cache/issues/166">#166</a>)</li> <li>See full diff in <a href="https://github.com/swatinem/rust-cache/compare/v2.6.2...v2.7.1">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=Swatinem/rust-cache&package-manager=github_actions&previous-version=2.6.2&new-version=2.7.1)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) You can trigger a rebase of this PR by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` show <dependency name> ignore conditions` will show all of the ignore conditions of the specified dependency - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) </details> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2023-11-02 10:48:25 +00:00
meili-bors[bot]	ea4a266f08	Merge #4182 4182: Bump mislav/bump-homebrew-formula-action from 2 to 3 r=curquiza a=dependabot[bot] Bumps [mislav/bump-homebrew-formula-action](https://github.com/mislav/bump-homebrew-formula-action) from 2 to 3. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/mislav/bump-homebrew-formula-action/releases">mislav/bump-homebrew-formula-action's releases</a>.</em></p> <blockquote> <h2>bump-homebrew-formula 3.0</h2> <h2>What's Changed</h2> <ul> <li>feat: bump to use node20 runtime by <a href="https://github.com/chenrui333"><code>`@chenrui333</code></a>` in <a href="https://redirect.github.com/mislav/bump-homebrew-formula-action/pull/61">mislav/bump-homebrew-formula-action#61</a></li> <li>Bump actions/checkout from 3 to 4 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/mislav/bump-homebrew-formula-action/pull/63">mislav/bump-homebrew-formula-action#63</a></li> <li>Bump <code>`@vercel/ncc</code>` from 0.34.0 to 0.38.0 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/mislav/bump-homebrew-formula-action/pull/67">mislav/bump-homebrew-formula-action#67</a></li> <li>Bump <code>`@actions/core</code>` from 1.9.1 to 1.10.1 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/mislav/bump-homebrew-formula-action/pull/68">mislav/bump-homebrew-formula-action#68</a></li> <li>Bump <code>`@octokit/core</code>` from 3.5.1 to 5.0.0 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/mislav/bump-homebrew-formula-action/pull/65">mislav/bump-homebrew-formula-action#65</a></li> <li>Bump TypeScript from 4.7 to 5.2</li> <li>Bump <code>`@typescript-eslint/eslint-plugin</code>` from 5.43.0 to 6.7.2 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/mislav/bump-homebrew-formula-action/pull/66">mislav/bump-homebrew-formula-action#66</a></li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/mislav/bump-homebrew-formula-action/compare/v2.4...v3.0">https://github.com/mislav/bump-homebrew-formula-action/compare/v2.4...v3.0</a></p> <h2>bump-homebrew-formula 2.4</h2> <h2>What's Changed</h2> <ul> <li>chore: use <code>/archive/refs/tags/${tagName}.tar.gz</code> rather than <code>/archive/${tagName}.tar.gz</code> by <a href="https://github.com/chenrui333"><code>`@chenrui333</code></a>` in <a href="https://redirect.github.com/mislav/bump-homebrew-formula-action/pull/53">mislav/bump-homebrew-formula-action#53</a></li> <li>Fix extracting version tags from GitHub download URLs by <a href="https://github.com/mislav"><code>`@mislav</code></a>` in <a href="https://redirect.github.com/mislav/bump-homebrew-formula-action/pull/62">mislav/bump-homebrew-formula-action#62</a></li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/mislav/bump-homebrew-formula-action/compare/v2.3...v2.4">https://github.com/mislav/bump-homebrew-formula-action/compare/v2.3...v2.4</a></p> <h2>bump-homebrew-formula 2.3</h2> <h2>What's Changed</h2> <ul> <li>Fix formula path after sharding of homebrew-core by <a href="https://github.com/williammartin"><code>`@williammartin</code></a>` in <a href="https://redirect.github.com/mislav/bump-homebrew-formula-action/pull/59">mislav/bump-homebrew-formula-action#59</a></li> <li>(docs): fix if condition in example by <a href="https://github.com/christian-bromann"><code>`@christian-bromann</code></a>` in <a href="https://redirect.github.com/mislav/bump-homebrew-formula-action/pull/54">mislav/bump-homebrew-formula-action#54</a></li> <li>(docs): use environment files instead of set-output by <a href="https://github.com/kyu08"><code>`@kyu08</code></a>` in <a href="https://redirect.github.com/mislav/bump-homebrew-formula-action/pull/57">mislav/bump-homebrew-formula-action#57</a></li> <li>Bump word-wrap from 1.2.3 to 1.2.4 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/mislav/bump-homebrew-formula-action/pull/55">mislav/bump-homebrew-formula-action#55</a></li> </ul> <h2>New Contributors</h2> <ul> <li><a href="https://github.com/christian-bromann"><code>`@christian-bromann</code></a>` made their first contribution in <a href="https://redirect.github.com/mislav/bump-homebrew-formula-action/pull/54">mislav/bump-homebrew-formula-action#54</a></li> <li><a href="https://github.com/kyu08"><code>`@kyu08</code></a>` made their first contribution in <a href="https://redirect.github.com/mislav/bump-homebrew-formula-action/pull/57">mislav/bump-homebrew-formula-action#57</a></li> <li><a href="https://github.com/williammartin"><code>`@williammartin</code></a>` made their first contribution in <a href="https://redirect.github.com/mislav/bump-homebrew-formula-action/pull/59">mislav/bump-homebrew-formula-action#59</a></li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/mislav/bump-homebrew-formula-action/compare/v2.2...v2.3">https://github.com/mislav/bump-homebrew-formula-action/compare/v2.2...v2.3</a></p> <h2>bump-homebrew-formula 2.2</h2> <h2>What's Changed</h2> <ul> <li>Fix scenario with generated GITHUB_TOKEN by <a href="https://github.com/mislav"><code>`@mislav</code></a>` in <a href="https://redirect.github.com/mislav/bump-homebrew-formula-action/pull/45">mislav/bump-homebrew-formula-action#45</a></li> <li>Bump <code>`@actions/core</code>` from 1.6.0 to 1.9.1 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/mislav/bump-homebrew-formula-action/pull/39">mislav/bump-homebrew-formula-action#39</a></li> <li>Bump minimatch from 3.0.4 to 3.1.2 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/mislav/bump-homebrew-formula-action/pull/40">mislav/bump-homebrew-formula-action#40</a></li> <li>Bump got and ava by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://redirect.github.com/mislav/bump-homebrew-formula-action/pull/41">mislav/bump-homebrew-formula-action#41</a></li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/mislav/bump-homebrew-formula-action/compare/v2.1...v2.2">https://github.com/mislav/bump-homebrew-formula-action/compare/v2.1...v2.2</a></p> <h2>bump-homebrew-formula 2.1</h2> <ul> <li>Fix extracting complex tag names from GitHub archive and release download URLs <a href="https://redirect.github.com/mislav/bump-homebrew-formula-action/issues/37">mislav/bump-homebrew-formula-action#37</a></li> </ul> </blockquote> </details> <details> <summary>Commits</summary> <ul> <li><a href="`b3327118b2`"><code>b332711</code></a> lib</li> <li><a href="`d1d8ac114e`"><code>d1d8ac1</code></a> Merge remote-tracking branch 'origin/main' into v3</li> <li><a href="`cf2d00157f`"><code>cf2d001</code></a> Fix calculating checksum for resource at download URL (<a href="https://redirect.github.com/mislav/bump-homebrew-formula-action/issues/77">#77</a>)</li> <li><a href="`2bcfdc9312`"><code>2bcfdc9</code></a> Merge pull request <a href="https://redirect.github.com/mislav/bump-homebrew-formula-action/issues/72">#72</a> from mislav/dependabot/npm_and_yarn/octokit/plugin-res...</li> <li><a href="`5678601dcb`"><code>5678601</code></a> Merge pull request <a href="https://redirect.github.com/mislav/bump-homebrew-formula-action/issues/74">#74</a> from mislav/dependabot/npm_and_yarn/eslint-8.50.0</li> <li><a href="`addc60eb43`"><code>addc60e</code></a> Bump <code>`@octokit/plugin-rest-endpoint-methods</code>` from 9.0.0 to 10.0.0</li> <li><a href="`44b3287225`"><code>44b3287</code></a> Merge pull request <a href="https://redirect.github.com/mislav/bump-homebrew-formula-action/issues/75">#75</a> from mislav/dependabot/npm_and_yarn/octokit/core-5.0.1</li> <li><a href="`fda81994d7`"><code>fda8199</code></a> Merge pull request <a href="https://redirect.github.com/mislav/bump-homebrew-formula-action/issues/71">#71</a> from mislav/dependabot/npm_and_yarn/octokit/request-er...</li> <li><a href="`2fd87fd7ea`"><code>2fd87fd</code></a> Bump <code>`@octokit/core</code>` from 5.0.0 to 5.0.1</li> <li><a href="`0c20930845`"><code>0c20930</code></a> Bump eslint from 8.49.0 to 8.50.0</li> <li>Additional commits viewable in <a href="https://github.com/mislav/bump-homebrew-formula-action/compare/v2...v3">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=mislav/bump-homebrew-formula-action&package-manager=github_actions&previous-version=2&new-version=3)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) You can trigger a rebase of this PR by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` show <dependency name> ignore conditions` will show all of the ignore conditions of the specified dependency - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) </details> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2023-11-02 08:48:19 +00:00
dependabot[bot]	49f069ed97	Bump Swatinem/rust-cache from 2.6.2 to 2.7.1 Bumps [Swatinem/rust-cache](https://github.com/swatinem/rust-cache) from 2.6.2 to 2.7.1. - [Release notes](https://github.com/swatinem/rust-cache/releases) - [Changelog](https://github.com/Swatinem/rust-cache/blob/master/CHANGELOG.md) - [Commits](https://github.com/swatinem/rust-cache/compare/v2.6.2...v2.7.1) --- updated-dependencies: - dependency-name: Swatinem/rust-cache dependency-type: direct:production update-type: version-update:semver-minor ... Signed-off-by: dependabot[bot] <support@github.com>	2023-11-01 17:57:42 +00:00
dependabot[bot]	be16b99d40	Bump actions/setup-node from 3 to 4 Bumps [actions/setup-node](https://github.com/actions/setup-node) from 3 to 4. - [Release notes](https://github.com/actions/setup-node/releases) - [Commits](https://github.com/actions/setup-node/compare/v3...v4) --- updated-dependencies: - dependency-name: actions/setup-node dependency-type: direct:production update-type: version-update:semver-major ... Signed-off-by: dependabot[bot] <support@github.com>	2023-11-01 17:57:38 +00:00
dependabot[bot]	ec0c09d17c	Bump docker/login-action from 2 to 3 Bumps [docker/login-action](https://github.com/docker/login-action) from 2 to 3. - [Release notes](https://github.com/docker/login-action/releases) - [Commits](https://github.com/docker/login-action/compare/v2...v3) --- updated-dependencies: - dependency-name: docker/login-action dependency-type: direct:production update-type: version-update:semver-major ... Signed-off-by: dependabot[bot] <support@github.com>	2023-11-01 17:57:33 +00:00
dependabot[bot]	a9230f6e6c	Bump mislav/bump-homebrew-formula-action from 2 to 3 Bumps [mislav/bump-homebrew-formula-action](https://github.com/mislav/bump-homebrew-formula-action) from 2 to 3. - [Release notes](https://github.com/mislav/bump-homebrew-formula-action/releases) - [Commits](https://github.com/mislav/bump-homebrew-formula-action/compare/v2...v3) --- updated-dependencies: - dependency-name: mislav/bump-homebrew-formula-action dependency-type: direct:production update-type: version-update:semver-major ... Signed-off-by: dependabot[bot] <support@github.com>	2023-11-01 17:57:30 +00:00
Clément Renault	b10c060bf7	Cleanup TOML	2023-11-01 14:03:04 +01:00
Clément Renault	e507ef5932	Slow the logging down	2023-11-01 13:49:32 +01:00
Clément Renault	c71b1d33ae	Sort entries using rayon in the transform sorters	2023-11-01 11:07:16 +01:00
Clément Renault	0fc446c62f	Add more timing logs to the Transform	2023-11-01 11:07:16 +01:00
Louis Dureuil	0fb6acefc3	Add snapshots for facets	2023-10-31 17:11:08 +01:00
Louis Dureuil	b1d1355b69	remove tests on soft-deleted	2023-10-31 16:36:27 +01:00
Louis Dureuil	f19332466e	Extract field value as values instead of Option<Value>	2023-10-31 16:36:27 +01:00
Louis Dureuil	03ddb4f310	use deladd in facet update tests	2023-10-31 16:36:27 +01:00
Louis Dureuil	c855cc2721	Remove unused test	2023-10-31 16:36:27 +01:00
Louis Dureuil	da0503ef80	Fix document count	2023-10-31 16:36:27 +01:00
meili-bors[bot]	54f0ee1ed2	Merge #4167 4167: Introduce the `meilitool` command line interface r=Kerollmops a=Kerollmops This PR introduces a small tool to help the Cloud team: - Clear the tasks queue by removing all the tasks - Dump a Meilisearch database without having to enqueue the task - Access this `meilitool` binary from the Docker Image ## TODO - [x] Modify the Docker File to ship with this new tool (`@curquiza,` could you review that, please?) - [x] Clear the tasks queue by removing all the tasks - [x] Add more logs to explain what is happening - [x] Clear the `update_files` folder - [x] Dump a Meilisearch database without having to enqueue the task - [x] Add more logs to explain what is happening - [x] Introduce a flag to skip dumping enqueued and processing tasks. - [x] Dump the instance uid. - [x] Dump the keys. - [x] Dump the tasks with the update files. - [x] Dump the index documents and settings. - [ ] ~Dump the experimental features~ Co-authored-by: Clément Renault <clement@meilisearch.com>	2023-10-31 14:05:22 +00:00
ManyTheFish	94206b0055	Update tests	2023-10-31 13:48:47 +01:00
Louis Dureuil	b40253bf18	update snapshots	2023-10-31 10:30:48 +01:00
Louis Dureuil	d8bf3f3fc2	Remove unused snapshots	2023-10-31 10:12:49 +01:00
Louis Dureuil	9d59e8011a	fix some tests	2023-10-31 10:08:36 +01:00
Louis Dureuil	dad78cbf8d	Bulk facet remove deletes keys from DB when value empty	2023-10-31 09:53:55 +01:00
Louis Dureuil	4e91707a06	Rename test	2023-10-31 09:41:17 +01:00
Louis Dureuil	de10f20732	Fix field distribution again	2023-10-30 17:47:22 +01:00
Clément Renault	ce5647e730	Fix Dockerfile WORKDIR path	2023-10-30 17:27:59 +01:00
Clément Renault	b57b818b67	Don't use the last version of clap	2023-10-30 16:57:31 +01:00
Clément Renault	f7ea94e5f4	Modify the Dockerfile to compile meilisearch and meilitool	2023-10-30 16:32:17 +01:00
Louis Dureuil	be395c7944	Change order of arguments to tokenizer_builder	2023-10-30 16:26:29 +01:00
Louis Dureuil	9fedd8101a	Fix tests	2023-10-30 15:11:07 +01:00
Louis Dureuil	54d07a8da3	Update field distribution taking into account both deletions and additions	2023-10-30 14:47:51 +01:00
Clément Renault	53382bb1b8	Introduce a new flag to skip dumping enqueued/processing tasks	2023-10-30 14:32:10 +01:00
Clément Renault	5b004a2583	Add more logs to the dump exporter	2023-10-30 14:31:55 +01:00
Clément Renault	13416ccbf7	Introduce a new meilitool to help the cloud team	2023-10-30 14:30:20 +01:00
Louis Dureuil	58690dfb19	Fix tests compilation after changes to ExternalDocumentsIds API	2023-10-30 13:34:07 +01:00
Louis Dureuil	abf424ebfc	Remove unused FromIterator	2023-10-30 11:41:56 +01:00
Clément Renault	dfab6293c9	Use an LMDB database to store the external documents ids	2023-10-30 11:41:23 +01:00
Louis Dureuil	fdf3f7f627	Fix facet distribution test	2023-10-30 11:41:23 +01:00
Louis Dureuil	6260cff65f	Actually delete documents from DB when the merge function says so	2023-10-30 11:41:22 +01:00
Louis Dureuil	8e0d9c9a5e	Recover delete_documents tests that were too eagerly deleted	2023-10-30 11:41:22 +01:00
Louis Dureuil	ae4ec8ea55	Add delete_document_using_wtxn to TempIndex	2023-10-30 11:41:22 +01:00
Louis Dureuil	652ac3052d	use new iterator in batch	2023-10-30 11:41:22 +01:00
Louis Dureuil	9a2dccc3bc	Add iterator to find external ids of a bitmap of internal ids	2023-10-30 11:41:22 +01:00
Louis Dureuil	a35988550c	Fix some snapshots	2023-10-30 11:41:22 +01:00
Louis Dureuil	e78281785c	Actually execute the transform even if there are only documents to delete	2023-10-30 11:41:22 +01:00
Louis Dureuil	3c15881818	Add simple delete test	2023-10-30 11:41:22 +01:00
Louis Dureuil	73c06d31d9	snapshot always display stuff in consistent order	2023-10-30 11:41:22 +01:00
Louis Dureuil	290e773d23	remove more warnings and fix some tests	2023-10-30 11:41:22 +01:00
Louis Dureuil	fa6c7f65ca	Add TmpIndex::delete_documents	2023-10-30 11:41:22 +01:00
Louis Dureuil	113527f466	Remove soft-deleted related methods from Index	2023-10-30 11:41:22 +01:00
Louis Dureuil	c534a1b687	Stop using delete documents pipeline in batch runner	2023-10-30 11:41:22 +01:00
Louis Dureuil	2263dff02b	Stop using removed delete pipelines almost everywhere	2023-10-30 11:41:22 +01:00
Louis Dureuil	d651b3ef01	Remove delete documents files	2023-10-30 11:41:20 +01:00
ManyTheFish	762b0b47e6	Use deladd merging function in chunks mergers	2023-10-30 11:40:20 +01:00
Louis Dureuil	01d5eedf2f	Remove some warnings	2023-10-30 11:40:20 +01:00
Louis Dureuil	073f89db79	Fix facet tests	2023-10-30 11:40:20 +01:00
Louis Dureuil	8370fbc92b	Fix snaps	2023-10-30 11:40:20 +01:00
Louis Dureuil	85f42fbc03	Handle external to internal id mapping from TypedChunk::Documents	2023-10-30 11:40:20 +01:00
Louis Dureuil	c6b3c18c85	WIP: Comment out document deletion in other pipelines than update TODO: fix calls to DELETE route	2023-10-30 11:40:20 +01:00
Louis Dureuil	bafeb892a7	Modify Index after changes to ExternalDocumentsIds	2023-10-30 11:40:20 +01:00
Louis Dureuil	8fb221dae3	Refactor ExternalDocumentsIds - Remove soft deleted - Add apply method that takes a list of operations to encapsulate modifications to the external -> internal mapping	2023-10-30 11:40:20 +01:00
Louis Dureuil	5be569e3e2	Update obkv	2023-10-30 11:40:20 +01:00
Louis Dureuil	946c762d28	WIP: reset documents in TypedChunk::Documents	2023-10-30 11:40:20 +01:00
Louis Dureuil	cda6ca1ee6	Remove TypedChunk::NewDocumentIds	2023-10-30 11:40:18 +01:00
Louis Dureuil	696fcf4d18	Fix document insertion into LMDB	2023-10-30 11:39:31 +01:00
ManyTheFish	476e4d3dbe	Use value buffer instead of the initial value when writting the final result in the sorter	2023-10-30 11:39:31 +01:00
Clément Renault	576fa9c6da	Remove useless comment	2023-10-30 11:39:31 +01:00
Kerollmops	77dcbff6b2	Remove and Insert the DelAdd geo points	2023-10-30 11:39:31 +01:00
Kerollmops	544440c363	Ignore geo fields when the Del and Add content is the same	2023-10-30 11:39:31 +01:00
Clément Renault	a3dae4db9b	Extract the geo fields DelAdd and generate a new DelAdd obkv with it	2023-10-30 11:39:31 +01:00
ManyTheFish	ba90a5ec0e	update extract fid word count docids	2023-10-30 11:39:31 +01:00
Louis Dureuil	b26dc9aabe	Explanatory code comment	2023-10-30 11:39:31 +01:00
Louis Dureuil	66abac9364	Use specialized `KvReaderDelAdd` type Co-authored-by: Clément Renault <clement@meilisearch.com>	2023-10-30 11:39:31 +01:00
Louis Dureuil	59f88c14b3	Simplify facet update after removing `Index::faceted_documents_ids`	2023-10-30 11:39:29 +01:00
Louis Dureuil	14832cb324	Remove Index::faceted_documents_ids	2023-10-30 11:37:32 +01:00
Louis Dureuil	04ec293024	Facet Incremental update	2023-10-30 11:37:30 +01:00
Louis Dureuil	f67ff3a738	Facets Bulk update	2023-10-30 11:36:40 +01:00
Clément Renault	560e8f5613	Introduce the CboRoaringBitmapCodec merge_deladd_into and use it	2023-10-30 11:34:55 +01:00
Clément Renault	2d3f15f82c	Introduce a function to only serialize the Add side of a DelAdd obkv	2023-10-30 11:34:55 +01:00
Clément Renault	40186bf403	Rename FieldIdWordCountDocids correctly	2023-10-30 11:34:50 +01:00
ManyTheFish	87e3d27878	update extract word pair proximity to support deladd obkvs	2023-10-30 11:34:02 +01:00
ManyTheFish	6bcf8b4f8c	update extract word position docids	2023-10-30 11:34:02 +01:00
ManyTheFish	46aa75abdb	update extract word docids	2023-10-30 11:34:02 +01:00
ManyTheFish	2597bbd107	Make script language docids map taking a tuple of roaring bitmaps expressing the deletions and the additions	2023-10-30 11:34:00 +01:00
Clément Renault	e2bc054604	Update extract_facet_string_docids to support deladd obkvs	2023-10-30 11:32:36 +01:00
Clément Renault	fcd3a1434d	Update extract_facet_number_docids to support deladd obkvs	2023-10-30 11:31:04 +01:00
Clément Renault	a82dee21e0	Rename docid_fid into fid_docid	2023-10-30 11:31:02 +01:00
Clément Renault	bc45c1206d	Implement all the facet extraction paths and simplify them	2023-10-30 11:29:08 +01:00
Clément Renault	6ae4100f07	Generate the DelAdd for is_null, is_empty, and exists	2023-10-30 11:29:08 +01:00
Clément Renault	0c47defeee	Work on fid docid facet values rewrite	2023-10-30 11:29:06 +01:00
ManyTheFish	313b16bec2	Support diff indexing on extract_docid_word_positions	2023-10-30 11:24:19 +01:00
ManyTheFish	1dd97578a8	Make the transform struct return diff-based documents obkvs	2023-10-30 11:22:07 +01:00
ManyTheFish	f5ef69293b	deactivate prefix dbs	2023-10-30 11:22:07 +01:00
ManyTheFish	1c5705c164	clean PR warnings	2023-10-30 11:22:05 +01:00
ManyTheFish	66c2c82a18	Split wpp in several sorters	2023-10-30 11:15:02 +01:00
ManyTheFish	28a8d0ccda	Fix word pair proximity	2023-10-30 11:15:02 +01:00
ManyTheFish	96be85396d	Use a vecDeque in wpp database	2023-10-30 11:15:02 +01:00
ManyTheFish	df9e5c8651	Generalize usage of CboRoaringBitmap codec to ease the use	2023-10-30 11:15:02 +01:00
ManyTheFish	b541d48847	Add buffer to the obkv writter	2023-10-30 11:15:02 +01:00
ManyTheFish	8ccf32d1a0	Compute word_fid_docids before word_docids and exact_word_docids	2023-10-30 11:15:02 +01:00
ManyTheFish	db1ca21231	add puffin in sorter into reeder function	2023-10-30 11:15:00 +01:00
ManyTheFish	11ea5acff9	Fix	2023-10-30 11:13:10 +01:00
ManyTheFish	8d77736a67	Fix fid_word_docids	2023-10-30 11:13:10 +01:00
ManyTheFish	748b333161	Add usefull debug assert before key insertion in database	2023-10-30 11:13:10 +01:00
ManyTheFish	17b647dfe5	Wip	2023-10-30 11:13:08 +01:00
meili-bors[bot]	2614e7d9ca	Merge #4174 4174: Fix warnings r=dureuill a=irevoire Fix all the warnings found in the CI: https://github.com/meilisearch/meilisearch/actions/runs/6622576021/job/17988323623 Co-authored-by: Tamo <tamo@meilisearch.com>	2023-10-30 10:12:54 +00:00
Tamo	e7244aa485	fix warnings	2023-10-30 11:00:46 +01:00
meili-bors[bot]	9cacc82307	Merge #4169 4169: update charabia r=curquiza a=ManyTheFish Update Charabia to v0.8.5 and add the new khmer tokenizer Co-authored-by: ManyTheFish <many@meilisearch.com>	2023-10-26 17:21:30 +00:00
ManyTheFish	4c6fddb1cb	update charabia	2023-10-26 17:01:10 +02:00
meili-bors[bot]	62ea81bef6	Merge #4132 4132: Extract the creation and last updated timestamp from v2 dumps r=irevoire a=vivek-26 # Pull Request ## Related issue Fixes #2989 ## What does this PR do? This PR - - extracts the `created_at` and `updated_at` dates from v2 dumps. - updates the unit tests. ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Vivek Kumar <vivek.26@outlook.com>	2023-10-24 08:50:57 +00:00
Vivek Kumar	f28f09ae2f	update tests for v2 dumps	2023-10-24 14:10:46 +05:30
meili-bors[bot]	ca52021079	Merge #4154 4154: Update version for the next release (v1.5.0) in Cargo.toml r=curquiza a=meili-bot ⚠️ This PR is automatically generated. Check the new version is the expected one and Cargo.lock has been updated before merging. Co-authored-by: curquiza <curquiza@users.noreply.github.com>	2023-10-23 12:00:50 +00:00
curquiza	ee6f79d60b	Update version for the next release (v1.5.0) in Cargo.toml	2023-10-23 11:49:07 +00:00
meili-bors[bot]	e4c24ca6a3	Merge #4151 4151: Bring back changes from v1.4.2 into `release-v1.5.0` r=dureuill a=curquiza This will bring the fixes in v1.4.2 for v1.5.0 release Co-authored-by: curquiza <curquiza@users.noreply.github.com> Co-authored-by: Vivek Kumar <vivek.26@outlook.com> Co-authored-by: Louis Dureuil <louis.dureuil@gmail.com>	2023-10-23 10:11:11 +00:00
Louis Dureuil	2bae9550c8	Add explanatory comment	2023-10-23 12:06:28 +02:00
Vivek Kumar	32c78ac8b1	add/update tests when search with distinct attribute & pagination with no ranking	2023-10-23 12:06:27 +02:00
Vivek Kumar	5fe7c4545a	compute all candidates correctly when skipping	2023-10-23 12:02:45 +02:00
curquiza	2042229927	Update version for the next release (v1.4.2) in Cargo.toml	2023-10-23 12:02:45 +02:00
meili-bors[bot]	eae9eab181	Merge #4126 4126: Make the experimental route /metrics activable via HTTP r=dureuill a=braddotcoffee # Pull Request ## Related issue Closes #4086 ## What does this PR do? - [x] Make `/metrics` available via HTTP as described in #4086 - [x] The users can still launch Meilisearch using the `--experimental-enable-metrics` flag. - [x] If the flag `--experimental-enable-metrics` is activated, a call to the `GET /experimental-features` route right after the launch will show `"metrics": true` even if the user has not called the `PATCH /experimental-features` route yet. - [x] Even if the --experimental-enable-metrics flag is present at launch, calling the `PATCH /experimental-features` route with `"metrics": false` disables the experimental feature. - [x] Update the spec - I was unable to find docs in this repository to update about the `/experimental-features` endpoint. I'll happily update if you point me in the right direction! ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Co-authored-by: bwbonanno <bradfordbonanno@gmail.com> Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-10-23 08:51:37 +00:00
Louis Dureuil	cf8dad1ca0	index_scheduler.features() is no longer fallible	2023-10-23 10:38:56 +02:00
bwbonanno	dd619913da	Use RwLock to never persist cli state to db	2023-10-19 12:45:57 -07:00
meili-bors[bot]	9b55ff16e9	Merge #4134 4134: Bump rustix from 0.36.15 to 0.36.16 r=Kerollmops a=dependabot[bot] Bumps [rustix](https://github.com/bytecodealliance/rustix) from 0.36.15 to 0.36.16. <details> <summary>Commits</summary> <ul> <li><a href="`6534992521`"><code>6534992</code></a> chore: Release rustix version 0.36.16</li> <li><a href="`4928cf7a38`"><code>4928cf7</code></a> Disable riscv64 testing.</li> <li><a href="`8cc159c4c3`"><code>8cc159c</code></a> Fix the <code>test_ttyname_ok</code> test when /dev/stdin is inaccessable. (<a href="https://redirect.github.com/bytecodealliance/rustix/issues/821">#821</a>)</li> <li><a href="`6dc7ba9478`"><code>6dc7ba9</code></a> Downgrade dependencies and disable tests to compile under Rust 1.48.</li> <li><a href="`ded8986e7e`"><code>ded8986</code></a> Disable MIPS in CI. (<a href="https://redirect.github.com/bytecodealliance/rustix/issues/793">#793</a>)</li> <li><a href="`739f9c3ba0`"><code>739f9c3</code></a> Fixes for <code>Dir</code> on macOS, FreeBSD, and WASI.</li> <li><a href="`87481a97f4`"><code>87481a9</code></a> Merge pull request from GHSA-c827-hfw6-qwvm</li> <li>See full diff in <a href="https://github.com/bytecodealliance/rustix/compare/v0.36.15...v0.36.16">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=rustix&package-manager=cargo&previous-version=0.36.15&new-version=0.36.16)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) Dependabot will resolve any conflicts with this PR as long as you don't alter it yourself. You can also trigger a rebase manually by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` show <dependency name> ignore conditions` will show all of the ignore conditions of the specified dependency - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) You can disable automated security fix PRs for this repo from the [Security Alerts page](https://github.com/meilisearch/meilisearch/network/alerts). </details> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2023-10-19 08:01:36 +00:00
dependabot[bot]	e761db582f	Bump rustix from 0.36.15 to 0.36.16 Bumps [rustix](https://github.com/bytecodealliance/rustix) from 0.36.15 to 0.36.16. - [Release notes](https://github.com/bytecodealliance/rustix/releases) - [Commits](https://github.com/bytecodealliance/rustix/compare/v0.36.15...v0.36.16) --- updated-dependencies: - dependency-name: rustix dependency-type: indirect ... Signed-off-by: dependabot[bot] <support@github.com>	2023-10-18 18:42:12 +00:00
bwbonanno	d8c649b3cd	Return recoverable error if we fail to retrieve metrics state	2023-10-18 08:28:24 -07:00
meili-bors[bot]	5e0485d8dd	Merge #4131 4131: Reduce proximity range from 7 to 3 r=Kerollmops a=ManyTheFish ## Summary This PR aims to reduce the impact of the proximity databases on the indexing time and on the database size by reducing the maximum distance between two words to be indexed in the proximity database. ## Stats ### Impact on database size and indexing time ![Impact on datasets](https://github.com/meilisearch/meilisearch/assets/6482087/28ed3d96-bdde-41c1-bdac-e90c1b1dbb23) ### Impact on search relevancy <details> \| dataset_name \| host_name \| Relevancy rate (Precision) \| completion_rate 25.00% \| completion_rate 50.00% \| completion_rate 75.00% \| completion_rate 100.00% \| \|--------------\|------------------\|------------------------------------\|-----------------\|-----------------\|-----------------\|-----------------\| \| FBIS \| 1_4_0 \| percentile-10 \| 0.00% \| 0.00% \| 0.00% \| 0.00% \| \| FBIS \| 1_4_0 \| percentile-25 \| 0.00% \| 0.00% \| 0.00% \| 0.00% \| \| FBIS \| 1_4_0 \| percentile-50 \| 0.00% \| 0.00% \| 5.00% \| 5.56% \| \| FBIS \| 1_4_0 \| percentile-75 \| 0.00% \| 12.50% \| 35.00% \| 45.00% \| \| FBIS \| 1_4_0 \| percentile-90 \| 20.00% \| 40.00% \| \| 100.00% \| \| FBIS \| 1_4_0 \| average \| 5.78% \| 11.16% \| 21.90% \| 26.29% \| \| FBIS \| reduce_proximity \| percentile-10 \| 0.00% \| 0.00% \| 0.00% \| 0.00% \| \| FBIS \| reduce_proximity \| percentile-25 \| 0.00% \| 0.00% \| 0.00% \| 0.00% \| \| FBIS \| reduce_proximity \| percentile-50 \| 0.00% \| 0.00% \| 5.00% \| 5.56% \| \| FBIS \| reduce_proximity \| percentile-75 \| 0.00% \| 15.00% \| 35.00% \| 40.00% \| \| FBIS \| reduce_proximity \| percentile-90 \| 20.00% \| 40.00% \| 85.00% \| 100.00% \| \| FBIS \| reduce_proximity \| average \| 5.55% \| 11.34% \| 21.75% \| 26.14% \| \| FR94 \| 1_4_0 \| percentile-10 \| 0.00% \| 0.00% \| 0.00% \| 0.00% \| \| FR94 \| 1_4_0 \| percentile-25 \| 0.00% \| 0.00% \| 0.00% \| 0.00% \| \| FR94 \| 1_4_0 \| percentile-50 \| 0.00% \| 0.00% \| 0.00% \| 0.00% \| \| FR94 \| 1_4_0 \| percentile-75 \| 0.00% \| 5.00% \| 15.00% \| 42.11% \| \| FR94 \| 1_4_0 \| percentile-90 \| 15.00% \| 54.55% \| 100.00% \| 100.00% \| \| FR94 \| 1_4_0 \| average \| 5.95% \| 12.07% \| 18.70% \| 25.57% \| \| FR94 \| reduce_proximity \| percentile-10 \| 0.00% \| 0.00% \| 0.00% \| 0.00% \| \| FR94 \| reduce_proximity \| percentile-25 \| 0.00% \| 0.00% \| 0.00% \| 0.00% \| \| FR94 \| reduce_proximity \| percentile-50 \| 0.00% \| 0.00% \| 0.00% \| 0.00% \| \| FR94 \| reduce_proximity \| percentile-75 \| 0.00% \| 5.00% \| 15.00% \| 42.11% \| \| FR94 \| reduce_proximity \| percentile-90 \| 15.00% \| 54.55% \| 100.00% \| 100.00% \| \| FR94 \| reduce_proximity \| average \| 5.79% \| 12.00% \| 18.70% \| 25.53% \| \| FT \| 1_4_0 \| percentile-10 \| 0.00% \| 0.00% \| 0.00% \| 0.00% \| \| FT \| 1_4_0 \| percentile-25 \| 0.00% \| 0.00% \| 0.00% \| 0.00% \| \| FT \| 1_4_0 \| percentile-50 \| 0.00% \| 0.00% \| 5.00% \| 10.00% \| \| FT \| 1_4_0 \| percentile-75 \| 0.00% \| 15.00% \| 30.00% \| 40.00% \| \| FT \| 1_4_0 \| percentile-90 \| 20.00% \| 50.00% \| 65.00% \| 100.00% \| \| FT \| 1_4_0 \| average \| 5.08% \| 12.58% \| 20.00% \| 25.49% \| \| FT \| reduce_proximity \| percentile-10 \| 0.00% \| 0.00% \| 0.00% \| 0.00% \| \| FT \| reduce_proximity \| percentile-25 \| 0.00% \| 0.00% \| 0.00% \| 0.00% \| \| FT \| reduce_proximity \| percentile-50 \| 0.00% \| 0.00% \| 5.00% \| 10.00% \| \| FT \| reduce_proximity \| percentile-75 \| 0.00% \| 15.00% \| 30.00% \| 40.00% \| \| FT \| reduce_proximity \| percentile-90 \| 10.00% \| 45.00% \| 60.00% \| 100.00% \| \| FT \| reduce_proximity \| average \| 5.01% \| 12.64% \| 20.10% \| 25.53% \| \| LAT \| 1_4_0 \| percentile-10 \| 0.00% \| 0.00% \| 0.00% \| 0.00% \| \| LAT \| 1_4_0 \| percentile-25 \| 0.00% \| 0.00% \| 0.00% \| 0.00% \| \| LAT \| 1_4_0 \| percentile-50 \| 0.00% \| 0.00% \| 5.00% \| 5.00% \| \| LAT \| 1_4_0 \| percentile-75 \| 5.00% \| 15.00% \| 30.00% \| 30.00% \| \| LAT \| 1_4_0 \| percentile-90 \| 15.00% \| 45.00% \| 60.00% \| 80.00% \| \| LAT \| 1_4_0 \| average \| 4.80% \| 11.80% \| 17.88% \| 21.62% \| \| LAT \| reduce_proximity \| percentile-10 \| 0.00% \| 0.00% \| 0.00% \| 0.00% \| \| LAT \| reduce_proximity \| percentile-25 \| 0.00% \| 0.00% \| 0.00% \| 0.00% \| \| LAT \| reduce_proximity \| percentile-50 \| 0.00% \| 0.00% \| 5.00% \| 5.00% \| \| LAT \| reduce_proximity \| percentile-75 \| 0.00% \| 11.11% \| 25.00% \| 35.00% \| \| LAT \| reduce_proximity \| percentile-90 \| 15.00% \| 45.00% \| 55.00% \| 80.00% \| \| LAT \| reduce_proximity \| average \| 4.43% \| 11.23% \| 17.32% \| 21.45% \| </details> ### Impact on Search time \| dataset_name \| host_name \| 25.00% \| 50.00% \| 75.00% \| 100.00% \| Average \| \|--------------\|------------------\|------------:\|------------:\|------------:\|------------:\|-------------\| \| FBIS \| 1_4_0 \| 3.45 \| 7.446666667 \| 9.773489933 \| 9.620300752 \| 7.572614338 \| \| FBIS \| reduce_proximity \| 2.983333333 \| 5.316666667 \| 6.911073826 \| 7.637218045 \| 5.712072968 \| \| FR94 \| 1_4_0 \| 2.236666667 \| 4.45 \| 5.523489933 \| 4.560150376 \| 4.192576744 \| \| FR94 \| reduce_proximity \| 2.09 \| 3.991666667 \| 4.981543624 \| 4.266917293 \| 3.832531896 \| \| FT \| 1_4_0 \| 5.956666667 \| 9.656666667 \| 13.86912752 \| 10.83270677 \| 10.0787919 \| \| FT \| reduce_proximity \| 4.51 \| 5.981666667 \| 7.701342282 \| 6.766917293 \| 6.23998156 \| \| LAT \| 1_4_0 \| 5.856666667 \| 9.233333333 \| 12.98322148 \| 10.78759398 \| 9.715203865 \| \| LAT \| reduce_proximity \| 6.91 \| 6.706666667 \| 8.463087248 \| 8.265037594 \| 7.586197877 \| ## Technical approach - Ensure the MAX_DISTANCE constant is used everywhere needed - Reduce the MAX_DISTANCE from 8 to 4 ## Related TBD Co-authored-by: ManyTheFish <many@meilisearch.com>	2023-10-18 14:56:08 +00:00
ManyTheFish	27eec21415	Fix tests	2023-10-18 16:03:22 +02:00
Vivek Kumar	62cc97ba70	update tests to include created_at and updated-at in v2 dumps	2023-10-18 13:31:39 +05:30
Vivek Kumar	fed59cc1d5	extract created_at and updated_at dates from v2 dumps	2023-10-18 13:30:24 +05:30
bwbonanno	2b3adef796	Use index_scheduler from configured app_data in middleware	2023-10-17 08:17:13 -07:00
bwbonanno	956cfc5487	Add runtime check to metrics middleware	2023-10-16 13:48:57 -07:00
bwbonanno	12fc878640	Merge remote-tracking branch 'origin/main' into enable-metrics-http	2023-10-16 13:48:01 -07:00
meili-bors[bot]	0a2e8b92a9	Merge #4129 4129: Add webinar banner in README r=curquiza a=curquiza Co-authored-by: curquiza <clementine@meilisearch.com>	2023-10-16 17:35:48 +00:00
meili-bors[bot]	c7a3f80de6	Merge #4073 4073: Simplify Puffin report exports r=ManyTheFish a=Kerollmops This PR changes how we export Puffin reports by directly writing them to disk when the `exportPuffinReports` [experimental feature is enabled](https://www.meilisearch.com/docs/learn/experimental/overview) on the `/experimental-features` route. It also adds more puffing logging to the deletion phase and grenad helpers. The puffin reports are identified by the date and time at which they are exported. ## Todo List - [x] Change the CLI flag to be an API experimental option. - [x] Create [a PRD for this experimental feature (private)](https://www.notion.so/meilisearch/Export-Puffin-Reports-091df151e71c4edfb7d72f4bf995b3ea). - [x] Create and complete [a product discussion](https://github.com/meilisearch/product/discussions/693) (copy/paste PROFILING markdown?). - [x] Update the _PROFILING.md_ markdown file instructions. - [x] Change the debug logs of the processing operation (visible in puffin viewer). Co-authored-by: Clément Renault <clement@meilisearch.com> Co-authored-by: Kerollmops <clement@meilisearch.com>	2023-10-16 15:48:15 +00:00
curquiza	029d4de043	Add webinar banner in README	2023-10-16 14:38:10 +02:00
meili-bors[bot]	549f1bcccf	Merge #4125 4125: Rename benchmark CI file to find it easily in the manifest list r=Kerollmops a=curquiza Co-authored-by: curquiza <clementine@meilisearch.com>	2023-10-16 11:38:28 +00:00
bwbonanno	689ec7c7ad	Make the experimental route /metrics activable via HTTP	2023-10-13 22:12:54 +00:00
Clément Renault	3655d4bdca	Move the puffin file export logic into the run function	2023-10-13 13:11:30 +02:00
Clément Renault	055ca3935b	Update index-scheduler/src/batch.rs Co-authored-by: Tamo <tamo@meilisearch.com>	2023-10-13 13:11:30 +02:00
Kerollmops	1b8871a585	Make cargo insta happy	2023-10-13 13:11:30 +02:00
Kerollmops	bf8fac6676	Fix the tests	2023-10-13 13:11:30 +02:00
Kerollmops	f2a9e1ebbb	Improve the debugging experience in the puffin reports	2023-10-13 13:11:30 +02:00
Kerollmops	c45c6cf54c	Update the PROFILING.md file	2023-10-13 13:11:30 +02:00
Kerollmops	513e61e9a3	Remove the experimental CLI flag	2023-10-13 13:11:29 +02:00
Kerollmops	90a626bf80	Use the runtime feature to enable puffin report exporting	2023-10-13 13:11:29 +02:00
Kerollmops	0d4acf2daa	Fix the metrics product URL	2023-10-13 13:11:29 +02:00
Kerollmops	58db8d85ec	Add the `exportPuffinReports` option to the runtime features route	2023-10-13 13:11:29 +02:00
Clément Renault	62dfd09dc6	Add more puffin logs to the deletion functions	2023-10-13 13:11:09 +02:00
Clément Renault	656dadabea	Expose an experimental flag to write the puffin reports to disk	2023-10-13 13:11:09 +02:00
Clément Renault	c5f7893fbb	Remove the puffin http dependency	2023-10-13 13:11:08 +02:00
curquiza	8cf2ccf168	Rename benchmark CI file to find it easily in the manifest list	2023-10-12 18:41:26 +02:00
meili-bors[bot]	0913373a5e	Merge #4122 4122: Bring back changes from `release-v1.4.1` into `main` r=Kerollmops a=curquiza Co-authored-by: curquiza <curquiza@users.noreply.github.com> Co-authored-by: meili-bors[bot] <89034592+meili-bors[bot]@users.noreply.github.com> Co-authored-by: Tamo <tamo@meilisearch.com> Co-authored-by: Vivek Kumar <vivek.26@outlook.com> Co-authored-by: Clément Renault <clement@meilisearch.com>	2023-10-12 15:57:47 +00:00
Clément Renault	1a7f1282af	Fix test to use new common Value type	2023-10-12 17:37:04 +02:00
Clément Renault	bc747aac3a	Cut the first 8 characters	2023-10-12 15:04:37 +02:00
Kerollmops	be92376ab3	Fix originating commit branch	2023-10-12 13:51:41 +02:00
Kerollmops	cf7e355735	Fix originating commit command	2023-10-12 13:12:53 +02:00
Clément Renault	5f09d89ad1	Fetch the whole git history when cloning	2023-10-12 12:25:26 +02:00
Clément Renault	6ecb26a3f8	Add more info on the commenting CI command	2023-10-12 11:54:56 +02:00
meili-bors[bot]	76c6f554d6	Merge #4101 4101: Bump webpki from 0.22.1 to 0.22.2 r=curquiza a=dependabot[bot] Bumps [webpki](https://github.com/briansmith/webpki) from 0.22.1 to 0.22.2. <details> <summary>Commits</summary> <ul> <li>See full diff in <a href="https://github.com/briansmith/webpki/commits">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=webpki&package-manager=cargo&previous-version=0.22.1&new-version=0.22.2)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) Dependabot will resolve any conflicts with this PR as long as you don't alter it yourself. You can also trigger a rebase manually by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` show <dependency name> ignore conditions` will show all of the ignore conditions of the specified dependency - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) You can disable automated security fix PRs for this repo from the [Security Alerts page](https://github.com/meilisearch/meilisearch/network/alerts). </details> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2023-10-12 08:46:04 +00:00
meili-bors[bot]	f343ef5f2f	Merge #4108 4108: Fix bug where search with distinct attribute and no ranking, returns offset+limit hits r=curquiza a=vivek-26 # Pull Request ## Related issue Fixes #4078 ## What does this PR do? This PR - - Fixes bug where search with distinct attribute and no ranking, returns offset+limit hits. - Adds unit and integration tests. ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Vivek Kumar <vivek.26@outlook.com>	2023-10-12 07:51:29 +00:00
Kerollmops	96982a768a	Triggers for every type of issue_comment	2023-10-11 23:18:29 +02:00
meili-bors[bot]	fca78fbc46	Merge #4082 4082: Update sprint_issue.md r=curquiza a=curquiza Following internal recent discussions Co-authored-by: Clémentine U. - curqui <clementine@meilisearch.com>	2023-10-11 15:12:38 +00:00
meili-bors[bot]	67a678cfb6	Merge #4089 4089: Use a bufreader and bufwriter everytime there is a grenad<file> r=curquiza a=irevoire # Pull Request Wrap all the files we give to a grenad in a `BufReader` or `BufWriter`. The dump import I tried in the issue went from 2h to 10 minutes on my machine. I also ran a bunch of benchmarks on my machine, and we're faster by a few seconds everywhere but nothing huge. ----- The one thing I’m afraid about is if we used to get the inner file in a grenad and then do a read right after without a seek at the beginning of the file or a reopen. Since we now use a bufreader our read would return the bytes one buffer later and probably completely corrupt what we were supposed to read. From what I see, it looks like it works, but I may have missed something, I don't know much about this part of the codebase. This issue should not arise on the bufwriter, though, because if we're not able to write the content of the buffer I ensured that the `into_inner` of the bufwriter should return an internal error. ## Related issue Fixes #4087 Co-authored-by: Tamo <tamo@meilisearch.com>	2023-10-11 14:27:00 +00:00
Vivek Kumar	d1331d8abf	add integration test for distinct search with no ranking	2023-10-11 19:12:56 +05:30
Vivek Kumar	19ba129165	add unit test for distinct search with no ranking	2023-10-11 19:02:27 +05:30
Vivek Kumar	d4da06ff47	fix bug where distinct search with no ranking returns offset+limit hits	2023-10-11 19:02:16 +05:30
Clément Renault	3e0471edae	Only trigger CI on created or edited comments	2023-10-11 15:15:15 +02:00
Clément Renault	432df03c4c	Use the correct base filename in the comment bench CI	2023-10-11 14:57:03 +02:00
Clément Renault	11958016dd	Force a small if to evoid triggering the CI every time	2023-10-11 14:27:51 +02:00
Kerollmops	63c250a04d	Do not use the GITHUB_REF variable	2023-10-11 13:05:54 +02:00
Clément Renault	06d8cd5b72	Make sure that we checkout on the right branch	2023-10-11 12:02:44 +02:00
Tamo	c0f2724c2d	get rids of the new introduced error code in favor of an io::Error	2023-10-10 15:12:23 +02:00
Tamo	d772073dfa	use a bufreader everytime there is a grenad<file>	2023-10-10 15:00:30 +02:00
meili-bors[bot]	8fe8ddea79	Merge #4112 4112: Update version for the next release (v1.4.1) in Cargo.toml r=curquiza a=meili-bot ⚠️ This PR is automatically generated. Check the new version is the expected one and Cargo.lock has been updated before merging. Co-authored-by: curquiza <curquiza@users.noreply.github.com>	2023-10-10 09:05:10 +00:00
curquiza	8a95bf28e5	Update version for the next release (v1.4.1) in Cargo.toml	2023-10-10 09:01:45 +00:00
Clément Renault	c0fd3dffb8	Setup a Github Token env var	2023-10-09 18:04:49 +02:00
Clément Renault	c42fd5375f	Fix the git commands again	2023-10-09 17:36:19 +02:00
Clément Renault	b418c3a756	Use the PAT token instead	2023-10-09 16:52:04 +02:00
Clément Renault	1cde455758	Fix workflow CI	2023-10-09 16:30:46 +02:00
Clément Renault	ca19bae72f	Prefer using a action to manage commands	2023-10-09 14:56:41 +02:00
meili-bors[bot]	705878ff59	Merge #4102 4102: Introduce the first bot that shows benchmarks results r=curquiza a=Kerollmops TBD Co-authored-by: Kerollmops <clement@meilisearch.com> Co-authored-by: Clément Renault <clement@meilisearch.com>	2023-10-05 10:11:06 +00:00
Clément Renault	92c280d1c8	Update .github/workflows/trigger-benchmarks-on-message.yml Co-authored-by: Clémentine U. - curqui <clementine@meilisearch.com>	2023-10-05 12:09:52 +02:00
Kerollmops	181e7a1e53	Introduce the first bot that triggers benchmarks	2023-10-05 12:05:38 +02:00
meili-bors[bot]	2e5abb4d2c	Merge #4098 4098: Bump docker/metadata-action from 4 to 5 r=curquiza a=dependabot[bot] Bumps [docker/metadata-action](https://github.com/docker/metadata-action) from 4 to 5. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/docker/metadata-action/releases">docker/metadata-action's releases</a>.</em></p> <blockquote> <h2>v5.0.0</h2> <ul> <li>Node 20 as default runtime (requires <a href="https://github.com/actions/runner/releases/tag/v2.308.0">Actions Runner v2.308.0</a> or later) by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` in <a href="https://redirect.github.com/docker/metadata-action/pull/328">docker/metadata-action#328</a></li> <li>Bump <code>`@actions/core</code>` from 1.10.0 to 1.10.1 in <a href="https://redirect.github.com/docker/metadata-action/pull/333">docker/metadata-action#333</a></li> <li>Bump csv-parse from 5.4.0 to 5.5.0 in <a href="https://redirect.github.com/docker/metadata-action/pull/320">docker/metadata-action#320</a></li> <li>Bump semver from 7.5.1 to 7.5.2 in <a href="https://redirect.github.com/docker/metadata-action/pull/304">docker/metadata-action#304</a></li> <li>Bump handlebars from 4.7.7 to 4.7.8 in <a href="https://redirect.github.com/docker/metadata-action/pull/315">docker/metadata-action#315</a></li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/metadata-action/compare/v4.6.0...v5.0.0">https://github.com/docker/metadata-action/compare/v4.6.0...v5.0.0</a></p> <h2>v4.6.0</h2> <ul> <li>Dedup and sort labels by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` in <a href="https://redirect.github.com/docker/metadata-action/pull/301">docker/metadata-action#301</a></li> <li>Bump <code>`@docker/actions-toolkit</code>` from 0.3.0 to 0.5.0 in <a href="https://redirect.github.com/docker/metadata-action/pull/302">docker/metadata-action#302</a></li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/metadata-action/compare/v4.5.0...v4.6.0">https://github.com/docker/metadata-action/compare/v4.5.0...v4.6.0</a></p> <h2>v4.5.0</h2> <ul> <li>Bump <code>`@docker/actions-toolkit</code>` from 0.1.0 to 0.3.0 in <a href="https://redirect.github.com/docker/metadata-action/pull/296">docker/metadata-action#296</a></li> <li>Bump csv-parse from 5.3.8 to 5.4.0 in <a href="https://redirect.github.com/docker/metadata-action/pull/294">docker/metadata-action#294</a></li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/metadata-action/compare/v4.4.0...v4.5.0">https://github.com/docker/metadata-action/compare/v4.4.0...v4.5.0</a></p> <h2>v4.4.0</h2> <ul> <li>Add <code>context</code> input to define the metadata provider by <a href="https://github.com/neilime"><code>`@neilime</code></a>` in <a href="https://redirect.github.com/docker/metadata-action/pull/248">docker/metadata-action#248</a></li> <li>Switch to actions-toolkit implementation by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` in <a href="https://redirect.github.com/docker/metadata-action/pull/266">docker/metadata-action#266</a> <a href="https://redirect.github.com/docker/metadata-action/pull/273">docker/metadata-action#273</a> <a href="https://redirect.github.com/docker/metadata-action/pull/284">docker/metadata-action#284</a></li> <li>Bump csv-parse from 5.3.3 to 5.3.8 in <a href="https://redirect.github.com/docker/metadata-action/pull/271">docker/metadata-action#271</a> <a href="https://redirect.github.com/docker/metadata-action/pull/286">docker/metadata-action#286</a></li> <li>Bump moment-timezone from 0.5.40 to 0.5.43 in <a href="https://redirect.github.com/docker/metadata-action/pull/268">docker/metadata-action#268</a> <a href="https://redirect.github.com/docker/metadata-action/pull/278">docker/metadata-action#278</a> <a href="https://redirect.github.com/docker/metadata-action/pull/281">docker/metadata-action#281</a></li> <li>Bump semver from 7.4.0 to 7.5.0 in <a href="https://redirect.github.com/docker/metadata-action/pull/285">docker/metadata-action#285</a></li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/metadata-action/compare/v4.3.0...v4.4.0">https://github.com/docker/metadata-action/compare/v4.3.0...v4.4.0</a></p> <h2>v4.3.0</h2> <ul> <li>Provide outputs as env vars by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://redirect.github.com/docker/metadata-action/issues/257">#257</a>)</li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/metadata-action/compare/v4.2.0...v4.3.0">https://github.com/docker/metadata-action/compare/v4.2.0...v4.3.0</a></p> <h2>v4.2.0</h2> <ul> <li>Add <code>tz</code> attribute to handlebar date function by <a href="https://github.com/chroju"><code>`@chroju</code></a>` (<a href="https://redirect.github.com/docker/metadata-action/issues/251">#251</a>)</li> <li>Bump minimatch from 3.0.4 to 3.1.2 (<a href="https://redirect.github.com/docker/metadata-action/issues/242">#242</a>)</li> <li>Bump csv-parse from 5.3.1 to 5.3.3 (<a href="https://redirect.github.com/docker/metadata-action/issues/245">#245</a>)</li> <li>Bump json5 from 2.2.0 to 2.2.3 (<a href="https://redirect.github.com/docker/metadata-action/issues/252">#252</a>)</li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/metadata-action/compare/v4.1.1...v4.2.0">https://github.com/docker/metadata-action/compare/v4.1.1...v4.2.0</a></p> <h2>v4.1.1</h2> <ul> <li>Revert changes to set associated head sha on pull request event by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://redirect.github.com/docker/metadata-action/issues/239">#239</a>) <ul> <li>User can still set associated head sha on PR by setting the env var <code>DOCKER_METADATA_PR_HEAD_SHA=true</code></li> </ul> </li> <li>Bump csv-parse from 5.3.0 to 5.3.1 (<a href="https://redirect.github.com/docker/metadata-action/issues/237">#237</a>)</li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/metadata-action/compare/v4.1.0...v4.1.1">https://github.com/docker/metadata-action/compare/v4.1.0...v4.1.1</a></p> <!-- raw HTML omitted --> </blockquote> <p>... (truncated)</p> </details> <details> <summary>Upgrade guide</summary> <p><em>Sourced from <a href="https://github.com/docker/metadata-action/blob/master/UPGRADE.md">docker/metadata-action's upgrade guide</a>.</em></p> <blockquote> <h1>Upgrade notes</h1> <h2>v2 to v3</h2> <ul> <li>Repository has been moved to docker org. Replace <code>crazy-max/ghaction-docker-meta@v2</code> with <code>docker/metadata-action@v5</code></li> <li>The default bake target has been changed: <code>ghaction-docker-meta</code> > <code>docker-metadata-action</code></li> </ul> <h2>v1 to v2</h2> <ul> <li><a href="https://github.com/docker/metadata-action/blob/master/#inputs">inputs</a> <ul> <li><a href="https://github.com/docker/metadata-action/blob/master/#tag-sha"><code>tag-sha</code></a></li> <li><a href="https://github.com/docker/metadata-action/blob/master/#tag-edge--tag-edge-branch"><code>tag-edge</code> / <code>tag-edge-branch</code></a></li> <li><a href="https://github.com/docker/metadata-action/blob/master/#tag-semver"><code>tag-semver</code></a></li> <li><a href="https://github.com/docker/metadata-action/blob/master/#tag-match--tag-match-group"><code>tag-match</code> / <code>tag-match-group</code></a></li> <li><a href="https://github.com/docker/metadata-action/blob/master/#tag-latest"><code>tag-latest</code></a></li> <li><a href="https://github.com/docker/metadata-action/blob/master/#tag-schedule"><code>tag-schedule</code></a></li> <li><a href="https://github.com/docker/metadata-action/blob/master/#tag-custom--tag-custom-only"><code>tag-custom</code> / <code>tag-custom-only</code></a></li> <li><a href="https://github.com/docker/metadata-action/blob/master/#label-custom"><code>label-custom</code></a></li> </ul> </li> <li><a href="https://github.com/docker/metadata-action/blob/master/#basic-workflow">Basic workflow</a></li> <li><a href="https://github.com/docker/metadata-action/blob/master/#semver-workflow">Semver workflow</a></li> </ul> <h3>inputs</h3> <table> <thead> <tr> <th>New</th> <th>Unchanged</th> <th>Removed</th> </tr> </thead> <tbody> <tr> <td><code>tags</code></td> <td><code>images</code></td> <td><code>tag-sha</code></td> </tr> <tr> <td><code>flavor</code></td> <td><code>sep-tags</code></td> <td><code>tag-edge</code></td> </tr> <tr> <td><code>labels</code></td> <td><code>sep-labels</code></td> <td><code>tag-edge-branch</code></td> </tr> <tr> <td></td> <td></td> <td><code>tag-semver</code></td> </tr> <tr> <td></td> <td></td> <td><code>tag-match</code></td> </tr> <tr> <td></td> <td></td> <td><code>tag-match-group</code></td> </tr> <tr> <td></td> <td></td> <td><code>tag-latest</code></td> </tr> <tr> <td></td> <td></td> <td><code>tag-schedule</code></td> </tr> <tr> <td></td> <td></td> <td><code>tag-custom</code></td> </tr> <tr> <td></td> <td></td> <td><code>tag-custom-only</code></td> </tr> <tr> <td></td> <td></td> <td><code>label-custom</code></td> </tr> </tbody> </table> <h4><code>tag-sha</code></h4> <pre lang="yaml"><code>tags: \| type=sha </code></pre> <h4><code>tag-edge</code> / <code>tag-edge-branch</code></h4> <pre lang="yaml"><code>tags: \| # default branch </tr></table> </code></pre> </blockquote> <p>... (truncated)</p> </details> <details> <summary>Commits</summary> <ul> <li><a href="`96383f4557`"><code>96383f4</code></a> Merge pull request <a href="https://redirect.github.com/docker/metadata-action/issues/320">#320</a> from docker/dependabot/npm_and_yarn/csv-parse-5.5.0</li> <li><a href="`f138b9677b`"><code>f138b96</code></a> chore: update generated content</li> <li><a href="`9cf7015b15`"><code>9cf7015</code></a> Bump csv-parse from 5.4.0 to 5.5.0</li> <li><a href="`5a8a5ff8df`"><code>5a8a5ff</code></a> Merge pull request <a href="https://redirect.github.com/docker/metadata-action/issues/315">#315</a> from docker/dependabot/npm_and_yarn/handlebars-4.7.8</li> <li><a href="`2279d9af58`"><code>2279d9a</code></a> chore: update generated content</li> <li><a href="`c659933213`"><code>c659933</code></a> Bump handlebars from 4.7.7 to 4.7.8</li> <li><a href="`48d23ccc05`"><code>48d23cc</code></a> Merge pull request <a href="https://redirect.github.com/docker/metadata-action/issues/333">#333</a> from docker/dependabot/npm_and_yarn/actions/core-1.10.1</li> <li><a href="`b83ffb48d6`"><code>b83ffb4</code></a> chore: update generated content</li> <li><a href="`3207f2405f`"><code>3207f24</code></a> Bump <code>`@actions/core</code>` from 1.10.0 to 1.10.1</li> <li><a href="`63f4a263e5`"><code>63f4a26</code></a> Merge pull request <a href="https://redirect.github.com/docker/metadata-action/issues/328">#328</a> from crazy-max/update-node20</li> <li>Additional commits viewable in <a href="https://github.com/docker/metadata-action/compare/v4...v5">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=docker/metadata-action&package-manager=github_actions&previous-version=4&new-version=5)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) You can trigger a rebase of this PR by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` show <dependency name> ignore conditions` will show all of the ignore conditions of the specified dependency - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) </details> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2023-10-04 08:40:29 +00:00
dependabot[bot]	44aaf5d9e3	Bump docker/metadata-action from 4 to 5 Bumps [docker/metadata-action](https://github.com/docker/metadata-action) from 4 to 5. - [Release notes](https://github.com/docker/metadata-action/releases) - [Upgrade guide](https://github.com/docker/metadata-action/blob/master/UPGRADE.md) - [Commits](https://github.com/docker/metadata-action/compare/v4...v5) --- updated-dependencies: - dependency-name: docker/metadata-action dependency-type: direct:production update-type: version-update:semver-major ... Signed-off-by: dependabot[bot] <support@github.com>	2023-10-03 23:09:04 +00:00
meili-bors[bot]	ff0ababf65	Merge #4097 4097: Bump docker/setup-buildx-action from 2 to 3 r=curquiza a=dependabot[bot] Bumps [docker/setup-buildx-action](https://github.com/docker/setup-buildx-action) from 2 to 3. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/docker/setup-buildx-action/releases">docker/setup-buildx-action's releases</a>.</em></p> <blockquote> <h2>v3.0.0</h2> <ul> <li>Node 20 as default runtime (requires <a href="https://github.com/actions/runner/releases/tag/v2.308.0">Actions Runner v2.308.0</a> or later) by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` in <a href="https://redirect.github.com/docker/setup-buildx-action/pull/264">docker/setup-buildx-action#264</a></li> <li>Bump <code>`@actions/core</code>` from 1.10.0 to 1.10.1 in <a href="https://redirect.github.com/docker/setup-buildx-action/pull/267">docker/setup-buildx-action#267</a></li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/setup-buildx-action/compare/v2.10.0...v3.0.0">https://github.com/docker/setup-buildx-action/compare/v2.10.0...v3.0.0</a></p> <h2>v2.10.0</h2> <h2>What's Changed</h2> <ul> <li>Bump <code>`@docker/actions-toolkit</code>` from 0.7.1 to 0.10.0 by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` in <a href="https://redirect.github.com/docker/setup-buildx-action/pull/258">docker/setup-buildx-action#258</a></li> <li>Bump word-wrap from 1.2.3 to 1.2.5 in <a href="https://redirect.github.com/docker/setup-buildx-action/pull/253">docker/setup-buildx-action#253</a></li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/setup-buildx-action/compare/v2.9.1...v2.10.0">https://github.com/docker/setup-buildx-action/compare/v2.9.1...v2.10.0</a></p> <h2>v2.9.1</h2> <ul> <li>Bump <code>`@docker/actions-toolkit</code>` from 0.7.0 to 0.7.1 in <a href="https://redirect.github.com/docker/setup-buildx-action/pull/248">docker/setup-buildx-action#248</a> <ul> <li>Fixes an issue where building Buildx does not match the local platform (<a href="https://redirect.github.com/docker/actions-toolkit/pull/135">docker/actions-toolkit#135</a>)</li> </ul> </li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/setup-buildx-action/compare/v2.9.0...v2.9.1">https://github.com/docker/setup-buildx-action/compare/v2.9.0...v2.9.1</a></p> <h2>v2.9.0</h2> <ul> <li>Bump <code>`@docker/actions-toolkit</code>` from 0.6.0 to 0.7.0 in <a href="https://redirect.github.com/docker/setup-buildx-action/pull/246">docker/setup-buildx-action#246</a> <ul> <li>Adds support to cache Buildx binary to hosted tool cache and GHA cache backend</li> </ul> </li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/setup-buildx-action/compare/v2.8.0...v2.9.0">https://github.com/docker/setup-buildx-action/compare/v2.8.0...v2.9.0</a></p> <h2>v2.8.0</h2> <ul> <li>Only set specific flags for drivers supporting them by <a href="https://github.com/nicks"><code>`@nicks</code></a>` in <a href="https://redirect.github.com/docker/setup-buildx-action/pull/241">docker/setup-buildx-action#241</a></li> <li>Bump <code>`@docker/actions-toolkit</code>` from 0.5.0 to 0.6.0 in <a href="https://redirect.github.com/docker/setup-buildx-action/pull/242">docker/setup-buildx-action#242</a></li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/setup-buildx-action/compare/v2.7.0...v2.8.0">https://github.com/docker/setup-buildx-action/compare/v2.7.0...v2.8.0</a></p> <h2>v2.7.0</h2> <ul> <li>Bump <code>`@docker/actions-toolkit</code>` from 0.3.0 to 0.5.0 in <a href="https://redirect.github.com/docker/setup-buildx-action/pull/237">docker/setup-buildx-action#237</a> <a href="https://redirect.github.com/docker/setup-buildx-action/pull/238">docker/setup-buildx-action#238</a></li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/setup-buildx-action/compare/v2.6.0...v2.7.0">https://github.com/docker/setup-buildx-action/compare/v2.6.0...v2.7.0</a></p> <h2>v2.6.0</h2> <ul> <li>Set node name for k8s driver when appending nodes by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` in <a href="https://redirect.github.com/docker/setup-buildx-action/pull/219">docker/setup-buildx-action#219</a></li> <li>Bump <code>`@docker/actions-toolkit</code>` from 0.1.0-beta.18 to 0.3.0 in <a href="https://redirect.github.com/docker/setup-buildx-action/pull/220">docker/setup-buildx-action#220</a> <a href="https://redirect.github.com/docker/setup-buildx-action/pull/229">docker/setup-buildx-action#229</a> <a href="https://redirect.github.com/docker/setup-buildx-action/pull/231">docker/setup-buildx-action#231</a> <a href="https://redirect.github.com/docker/setup-buildx-action/pull/236">docker/setup-buildx-action#236</a></li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/setup-buildx-action/compare/v2.5.0...v2.6.0">https://github.com/docker/setup-buildx-action/compare/v2.5.0...v2.6.0</a></p> <h2>v2.5.0</h2> <ul> <li><code>cleanup</code> input to remove builder and temp files by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` in <a href="https://redirect.github.com/docker/setup-buildx-action/pull/213">docker/setup-buildx-action#213</a></li> <li>do not remove builder using the <code>docker</code> driver by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` in <a href="https://redirect.github.com/docker/setup-buildx-action/pull/218">docker/setup-buildx-action#218</a></li> <li>fix current context as builder name for <code>docker</code> driver by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` in <a href="https://redirect.github.com/docker/setup-buildx-action/pull/209">docker/setup-buildx-action#209</a></li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/setup-buildx-action/compare/v2.4.1...v2.5.0">https://github.com/docker/setup-buildx-action/compare/v2.4.1...v2.5.0</a></p> <h2>v2.4.1</h2> <!-- raw HTML omitted --> </blockquote> <p>... (truncated)</p> </details> <details> <summary>Commits</summary> <ul> <li><a href="`f95db51fdd`"><code>f95db51</code></a> Merge pull request <a href="https://redirect.github.com/docker/setup-buildx-action/issues/267">#267</a> from docker/dependabot/npm_and_yarn/actions/core-1.10.1</li> <li><a href="`998a87c2c1`"><code>998a87c</code></a> chore: update generated content</li> <li><a href="`28bae59336`"><code>28bae59</code></a> build(deps): bump <code>`@actions/core</code>` from 1.10.0 to 1.10.1</li> <li><a href="`c215341715`"><code>c215341</code></a> Merge pull request <a href="https://redirect.github.com/docker/setup-buildx-action/issues/264">#264</a> from crazy-max/update-node20</li> <li><a href="`02e9319239`"><code>02e9319</code></a> chore: node 20 as default runtime</li> <li><a href="`5c9160effc`"><code>5c9160e</code></a> chore: update generated content</li> <li><a href="`1283140f57`"><code>1283140</code></a> chore: fix author in package.json</li> <li><a href="`c6afe06e4a`"><code>c6afe06</code></a> vendor: bump <code>`@docker/actions-toolkit</code>` from 0.10.0 to 0.12.0</li> <li><a href="`f35e0d5a04`"><code>f35e0d5</code></a> chore: update dev dependencies</li> <li><a href="`baeb468fb2`"><code>baeb468</code></a> dev: remove unneeded binaries</li> <li>Additional commits viewable in <a href="https://github.com/docker/setup-buildx-action/compare/v2...v3">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=docker/setup-buildx-action&package-manager=github_actions&previous-version=2&new-version=3)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) You can trigger a rebase of this PR by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` show <dependency name> ignore conditions` will show all of the ignore conditions of the specified dependency - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) </details> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2023-10-03 17:14:34 +00:00
dependabot[bot]	c5336af1c5	Bump docker/setup-buildx-action from 2 to 3 Bumps [docker/setup-buildx-action](https://github.com/docker/setup-buildx-action) from 2 to 3. - [Release notes](https://github.com/docker/setup-buildx-action/releases) - [Commits](https://github.com/docker/setup-buildx-action/compare/v2...v3) --- updated-dependencies: - dependency-name: docker/setup-buildx-action dependency-type: direct:production update-type: version-update:semver-major ... Signed-off-by: dependabot[bot] <support@github.com>	2023-10-03 15:41:06 +00:00
meili-bors[bot]	1567758a56	Merge #4099 4099: Bump docker/build-push-action from 4 to 5 r=curquiza a=dependabot[bot] Bumps [docker/build-push-action](https://github.com/docker/build-push-action) from 4 to 5. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/docker/build-push-action/releases">docker/build-push-action's releases</a>.</em></p> <blockquote> <h2>v5.0.0</h2> <ul> <li>Node 20 as default runtime (requires <a href="https://github.com/actions/runner/releases/tag/v2.308.0">Actions Runner v2.308.0</a> or later) by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` in <a href="https://redirect.github.com/docker/build-push-action/pull/954">docker/build-push-action#954</a></li> <li>Bump <code>`@actions/core</code>` from 1.10.0 to 1.10.1 in <a href="https://redirect.github.com/docker/build-push-action/pull/959">docker/build-push-action#959</a></li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/build-push-action/compare/v4.2.1...v5.0.0">https://github.com/docker/build-push-action/compare/v4.2.1...v5.0.0</a></p> <h2>v4.2.1</h2> <blockquote> <p><strong>Note</strong></p> <p>Buildx v0.10 enables support for a minimal <a href="https://slsa.dev/provenance/">SLSA Provenance</a> attestation, which requires support for <a href="https://github.com/opencontainers/image-spec">OCI-compliant</a> multi-platform images. This may introduce issues with registry and runtime support (e.g. <a href="https://redirect.github.com/docker/buildx/issues/1533">Google Cloud Run and AWS Lambda</a>). You can optionally disable the default provenance attestation functionality using <code>provenance: false</code>.</p> </blockquote> <ul> <li>warn if docker config can't be parsed by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` in <a href="https://redirect.github.com/docker/build-push-action/pull/957">docker/build-push-action#957</a></li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/build-push-action/compare/v4.2.0...v4.2.1">https://github.com/docker/build-push-action/compare/v4.2.0...v4.2.1</a></p> <h2>v4.2.0</h2> <blockquote> <p><strong>Note</strong></p> <p>Buildx v0.10 enables support for a minimal <a href="https://slsa.dev/provenance/">SLSA Provenance</a> attestation, which requires support for <a href="https://github.com/opencontainers/image-spec">OCI-compliant</a> multi-platform images. This may introduce issues with registry and runtime support (e.g. <a href="https://redirect.github.com/docker/buildx/issues/1533">Google Cloud Run and AWS Lambda</a>). You can optionally disable the default provenance attestation functionality using <code>provenance: false</code>.</p> </blockquote> <ul> <li>display proxy configuration by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` in <a href="https://redirect.github.com/docker/build-push-action/pull/872">docker/build-push-action#872</a></li> <li>chore(deps): Bump <code>`@docker/actions-toolkit</code>` from 0.6.0 to 0.8.0 in <a href="https://redirect.github.com/docker/build-push-action/pull/930">docker/build-push-action#930</a></li> <li>chore(deps): Bump word-wrap from 1.2.3 to 1.2.5 in <a href="https://redirect.github.com/docker/build-push-action/pull/925">docker/build-push-action#925</a></li> <li>chore(deps): Bump semver from 6.3.0 to 6.3.1 in <a href="https://redirect.github.com/docker/build-push-action/pull/902">docker/build-push-action#902</a></li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/build-push-action/compare/v4.1.1...v4.2.0">https://github.com/docker/build-push-action/compare/v4.1.1...v4.2.0</a></p> <h2>v4.1.1</h2> <blockquote> <p><strong>Note</strong></p> <p>Buildx v0.10 enables support for a minimal <a href="https://slsa.dev/provenance/">SLSA Provenance</a> attestation, which requires support for <a href="https://github.com/opencontainers/image-spec">OCI-compliant</a> multi-platform images. This may introduce issues with registry and runtime support (e.g. <a href="https://redirect.github.com/docker/buildx/issues/1533">Google Cloud Run and AWS Lambda</a>). You can optionally disable the default provenance attestation functionality using <code>provenance: false</code>.</p> </blockquote> <ul> <li>Bump <code>`@docker/actions-toolkit</code>` from 0.3.0 to 0.5.0 by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` in <a href="https://redirect.github.com/docker/build-push-action/pull/880">docker/build-push-action#880</a></li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/build-push-action/compare/v4.1.0...v4.1.1">https://github.com/docker/build-push-action/compare/v4.1.0...v4.1.1</a></p> <h2>v4.1.0</h2> <blockquote> <p><strong>Note</strong></p> <p>Buildx v0.10 enables support for a minimal <a href="https://slsa.dev/provenance/">SLSA Provenance</a> attestation, which requires support for <a href="https://github.com/opencontainers/image-spec">OCI-compliant</a> multi-platform images. This may introduce issues with registry and runtime support (e.g. <a href="https://redirect.github.com/docker/buildx/issues/1533">Google Cloud Run and AWS Lambda</a>). You can optionally disable the default provenance attestation functionality using <code>provenance: false</code>.</p> </blockquote> <ul> <li>Switch to actions-toolkit implementation by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` in <a href="https://redirect.github.com/docker/build-push-action/pull/811">docker/build-push-action#811</a> <a href="https://redirect.github.com/docker/build-push-action/pull/838">docker/build-push-action#838</a> <a href="https://redirect.github.com/docker/build-push-action/pull/855">docker/build-push-action#855</a> <a href="https://redirect.github.com/docker/build-push-action/pull/860">docker/build-push-action#860</a> <a href="https://redirect.github.com/docker/build-push-action/pull/875">docker/build-push-action#875</a></li> <li>e2e: quay.io by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` in <a href="https://redirect.github.com/docker/build-push-action/pull/799">docker/build-push-action#799</a> <a href="https://redirect.github.com/docker/build-push-action/pull/805">docker/build-push-action#805</a></li> <li>e2e: local harbor and nexus by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` in <a href="https://redirect.github.com/docker/build-push-action/pull/800">docker/build-push-action#800</a></li> <li>e2e: add artifactory container registry to test against by <a href="https://github.com/jedevc"><code>`@jedevc</code></a>` in <a href="https://redirect.github.com/docker/build-push-action/pull/804">docker/build-push-action#804</a></li> <li>e2e: add distribution tests by <a href="https://github.com/jedevc"><code>`@jedevc</code></a>` in <a href="https://redirect.github.com/docker/build-push-action/pull/814">docker/build-push-action#814</a> <a href="https://redirect.github.com/docker/build-push-action/pull/815">docker/build-push-action#815</a></li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/build-push-action/compare/v4.0.0...v4.1.0">https://github.com/docker/build-push-action/compare/v4.0.0...v4.1.0</a></p> </blockquote> </details> <details> <summary>Commits</summary> <ul> <li><a href="`0565240e2d`"><code>0565240</code></a> Merge pull request <a href="https://redirect.github.com/docker/build-push-action/issues/959">#959</a> from docker/dependabot/npm_and_yarn/actions/core-1.10.1</li> <li><a href="`3ab07f8801`"><code>3ab07f8</code></a> chore: update generated content</li> <li><a href="`b9e7e4daec`"><code>b9e7e4d</code></a> chore(deps): Bump <code>`@actions/core</code>` from 1.10.0 to 1.10.1</li> <li><a href="`04d1a3b049`"><code>04d1a3b</code></a> Merge pull request <a href="https://redirect.github.com/docker/build-push-action/issues/954">#954</a> from crazy-max/update-node20</li> <li><a href="`1a4d1a13fb`"><code>1a4d1a1</code></a> chore: node 20 as default runtime</li> <li><a href="`675965c0e1`"><code>675965c</code></a> chore: update generated content</li> <li><a href="`58ee34cb6b`"><code>58ee34c</code></a> chore: fix author in package.json</li> <li><a href="`c97c4060bd`"><code>c97c406</code></a> fix ProxyConfig type when checking length</li> <li><a href="`47d5369e0b`"><code>47d5369</code></a> vendor: bump <code>`@docker/actions-toolkit</code>` from 0.8.0 to 0.12.0</li> <li><a href="`8895c7468f`"><code>8895c74</code></a> chore: update dev dependencies</li> <li>Additional commits viewable in <a href="https://github.com/docker/build-push-action/compare/v4...v5">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=docker/build-push-action&package-manager=github_actions&previous-version=4&new-version=5)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) You can trigger a rebase of this PR by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` show <dependency name> ignore conditions` will show all of the ignore conditions of the specified dependency - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) </details> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2023-10-03 11:55:31 +00:00
dependabot[bot]	37953afe1a	Bump docker/build-push-action from 4 to 5 Bumps [docker/build-push-action](https://github.com/docker/build-push-action) from 4 to 5. - [Release notes](https://github.com/docker/build-push-action/releases) - [Commits](https://github.com/docker/build-push-action/compare/v4...v5) --- updated-dependencies: - dependency-name: docker/build-push-action dependency-type: direct:production update-type: version-update:semver-major ... Signed-off-by: dependabot[bot] <support@github.com>	2023-10-03 11:53:53 +00:00
ManyTheFish	43989fe2e4	Reduce porximity range from 7 to 3	2023-10-03 12:16:48 +02:00
meili-bors[bot]	de3f992ae4	Merge #4095 4095: Bump docker/setup-qemu-action from 2 to 3 r=curquiza a=dependabot[bot] Bumps [docker/setup-qemu-action](https://github.com/docker/setup-qemu-action) from 2 to 3. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/docker/setup-qemu-action/releases">docker/setup-qemu-action's releases</a>.</em></p> <blockquote> <h2>v3.0.0</h2> <ul> <li>Node 20 as default runtime (requires <a href="https://github.com/actions/runner/releases/tag/v2.308.0">Actions Runner v2.308.0</a> or later) by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` in <a href="https://redirect.github.com/docker/setup-qemu-action/pull/102">docker/setup-qemu-action#102</a></li> <li>Bump <code>`@actions/core</code>` from 1.10.0 to 1.10.1 in <a href="https://redirect.github.com/docker/setup-qemu-action/pull/103">docker/setup-qemu-action#103</a></li> <li>Bump semver from 6.3.0 to 6.3.1 in <a href="https://redirect.github.com/docker/setup-qemu-action/pull/89">docker/setup-qemu-action#89</a></li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/setup-qemu-action/compare/v2.2.0...v3.0.0">https://github.com/docker/setup-qemu-action/compare/v2.2.0...v3.0.0</a></p> <h2>v2.2.0</h2> <ul> <li>Trim off spaces in <code>platforms</code> input by <a href="https://github.com/Chocobo1"><code>`@Chocobo1</code></a>` in <a href="https://redirect.github.com/docker/setup-qemu-action/pull/64">docker/setup-qemu-action#64</a></li> <li>Switch to actions-toolkit implementation by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` in <a href="https://redirect.github.com/docker/setup-qemu-action/pull/70">docker/setup-qemu-action#70</a> <a href="https://redirect.github.com/docker/setup-qemu-action/pull/80">docker/setup-qemu-action#80</a> <a href="https://redirect.github.com/docker/setup-qemu-action/pull/83">docker/setup-qemu-action#83</a></li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/setup-qemu-action/compare/v2.1.0...v2.2.0">https://github.com/docker/setup-qemu-action/compare/v2.1.0...v2.2.0</a></p> <h2>v2.1.0</h2> <ul> <li>Use context for inputs by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://redirect.github.com/docker/setup-qemu-action/issues/62">#62</a>)</li> <li>Use built-in <code>getExecOutput</code> by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://redirect.github.com/docker/setup-qemu-action/issues/61">#61</a>)</li> <li>Remove workaround for <code>setOutput</code> by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://redirect.github.com/docker/setup-qemu-action/issues/63">#63</a>)</li> <li>Bump <code>`@actions/core</code>` from 1.6.0 to 1.10.0 (<a href="https://redirect.github.com/docker/setup-qemu-action/issues/54">#54</a> <a href="https://redirect.github.com/docker/setup-qemu-action/issues/58">#58</a> <a href="https://redirect.github.com/docker/setup-qemu-action/issues/59">#59</a>)</li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/setup-qemu-action/compare/v2.0.0...v2.1.0">https://github.com/docker/setup-qemu-action/compare/v2.0.0...v2.1.0</a></p> </blockquote> </details> <details> <summary>Commits</summary> <ul> <li><a href="`68827325e0`"><code>6882732</code></a> Merge pull request <a href="https://redirect.github.com/docker/setup-qemu-action/issues/103">#103</a> from docker/dependabot/npm_and_yarn/actions/core-1.10.1</li> <li><a href="`183f4af504`"><code>183f4af</code></a> chore: update generated content</li> <li><a href="`f17493529e`"><code>f174935</code></a> build(deps): bump <code>`@actions/core</code>` from 1.10.0 to 1.10.1</li> <li><a href="`2e423eb500`"><code>2e423eb</code></a> Merge pull request <a href="https://redirect.github.com/docker/setup-qemu-action/issues/89">#89</a> from docker/dependabot/npm_and_yarn/semver-6.3.1</li> <li><a href="`ecc406afa7`"><code>ecc406a</code></a> Bump semver from 6.3.0 to 6.3.1</li> <li><a href="`12dec5e201`"><code>12dec5e</code></a> Merge pull request <a href="https://redirect.github.com/docker/setup-qemu-action/issues/102">#102</a> from crazy-max/update-node20</li> <li><a href="`c29b312130`"><code>c29b312</code></a> chore: node 20 as default runtime</li> <li><a href="`34ae628c8f`"><code>34ae628</code></a> chore: update generated content</li> <li><a href="`1f3d2e1ac0`"><code>1f3d2e1</code></a> chore: fix author in package.json</li> <li><a href="`277dbe8c9c`"><code>277dbe8</code></a> vendor: bump <code>`@docker/actions-toolkit</code>` from 0.3.0 to 0.12.0</li> <li>Additional commits viewable in <a href="https://github.com/docker/setup-qemu-action/compare/v2...v3">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=docker/setup-qemu-action&package-manager=github_actions&previous-version=2&new-version=3)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) You can trigger a rebase of this PR by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` show <dependency name> ignore conditions` will show all of the ignore conditions of the specified dependency - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) </details> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2023-10-03 08:20:50 +00:00
dependabot[bot]	c668a29ed5	Bump webpki from 0.22.1 to 0.22.2 Bumps [webpki](https://github.com/briansmith/webpki) from 0.22.1 to 0.22.2. - [Commits](https://github.com/briansmith/webpki/commits) --- updated-dependencies: - dependency-name: webpki dependency-type: indirect ... Signed-off-by: dependabot[bot] <support@github.com>	2023-10-02 21:53:45 +00:00
dependabot[bot]	98f0618065	Bump docker/setup-qemu-action from 2 to 3 Bumps [docker/setup-qemu-action](https://github.com/docker/setup-qemu-action) from 2 to 3. - [Release notes](https://github.com/docker/setup-qemu-action/releases) - [Commits](https://github.com/docker/setup-qemu-action/compare/v2...v3) --- updated-dependencies: - dependency-name: docker/setup-qemu-action dependency-type: direct:production update-type: version-update:semver-major ... Signed-off-by: dependabot[bot] <support@github.com>	2023-10-01 17:18:20 +00:00
Clémentine U. - curqui	b10eeb0e41	Update .github/ISSUE_TEMPLATE/sprint_issue.md	2023-09-26 16:47:04 +02:00
Clémentine U. - curqui	4a8515e9fc	Update sprint_issue.md	2023-09-26 16:46:18 +02:00
meili-bors[bot]	86b314626d	Merge #4080 4080: Bring back changes from v1.4.0 into main r=Kerollmops a=curquiza Co-authored-by: ManyTheFish <many@meilisearch.com> Co-authored-by: Clément Renault <clement@meilisearch.com> Co-authored-by: Kerollmops <clement@meilisearch.com> Co-authored-by: meili-bors[bot] <89034592+meili-bors[bot]@users.noreply.github.com> Co-authored-by: curquiza <curquiza@users.noreply.github.com> Co-authored-by: Tamo <tamo@meilisearch.com> Co-authored-by: curquiza <clementine@meilisearch.com> Co-authored-by: Vivek Kumar <vivek.26@outlook.com> Co-authored-by: dogukanakkaya <doguakkaya27@hotmail.com>	2023-09-26 08:13:49 +00:00
meili-bors[bot]	bb79bdb3f8	Merge #4074 4074: Enable analytics in debug builds r=Kerollmops a=irevoire # Pull Request ## Related issue Fixes https://github.com/meilisearch/meilisearch/issues/4072 ## What does this PR do? - Stop disabling the analytics if meilisearch has been compiled in debug mode Co-authored-by: Tamo <tamo@meilisearch.com>	2023-09-21 15:54:41 +00:00
Tamo	d429e7da99	make clippy happy	2023-09-21 17:41:12 +02:00
Tamo	584b772248	enable metrics in debug builds	2023-09-21 17:01:05 +02:00
meili-bors[bot]	1806c04a9a	Merge #4065 4065: Dependency issue every 6 months r=curquiza a=curquiza To avoid spending too much time on it (1 every two sprints) If you disagree `@Kerollmops,` for security or any reason, please close the PR Co-authored-by: Clémentine U. - curqui <clementine@meilisearch.com>	2023-09-19 15:58:06 +00:00
Clémentine U. - curqui	3485e8f1c4	Update .github/workflows/dependency-issue.yml	2023-09-18 09:59:22 +02:00
Clémentine U. - curqui	fe697a6685	Dependency issue every 6 months	2023-09-18 09:57:58 +02:00
meili-bors[bot]	eb4135f8ae	Merge #4044 4044: Add more integrations to SDK CI r=curquiza a=curquiza For the integration scope management, but also to anticipate bugs and breaking changes for engine team, we need to add more SDKs tests into the CI Co-authored-by: curquiza <clementine@meilisearch.com>	2023-09-13 14:05:41 +00:00
curquiza	ec4844c3a6	Add dart, swift, dotnet, and java test Display docker image Add strapi and firebase Add rails and symfony tests Remove strapi and firestore tests Fix dotnet SDK CI Use specific dart SDK version Disable coverage for ruby SDK Prevent pushing coverage information to codecov Remove codecoverage token Trigger Build Trigger Build Trigger Build Trigger Build Trigger Build	2023-09-12 17:31:17 +02:00
meili-bors[bot]	77c3787b78	Merge #4056 4056: Rewrite segment_analytics module with the destructuring syntax r=Kerollmops a=vivek-26 # Pull Request ## Related issue Fixes #3928 ## What does this PR do? - This PR uses Rust Destructuring syntax in the `segment_analytics` module, such that adding or deleting fields causes an error at compile time. ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Vivek Kumar <vivek.26@outlook.com>	2023-09-12 08:07:50 +00:00
Vivek Kumar	4f902490b9	struct destructuring for DocumentsFetchAggregator	2023-09-12 10:39:28 +05:30
Vivek Kumar	1faee92748	struct destructuring for HealthAggregator	2023-09-12 10:39:28 +05:30
Vivek Kumar	5831466525	struct destructuring for DocumentsDeletionAggregator and TasksAggregator	2023-09-12 10:39:28 +05:30
Vivek Kumar	3cdb3e4eaf	struct destructuring for DocumentsAggregator	2023-09-12 10:39:27 +05:30
Vivek Kumar	26f34ec7a2	struct destructuring for FacetSearchAggregator	2023-09-12 10:39:27 +05:30
Vivek Kumar	07d36180ad	struct destructuring for MultiSearchAggregator	2023-09-12 10:39:27 +05:30
Vivek Kumar	4c641b79a2	use rust struct destructuring for SearchAggregator	2023-09-12 10:39:27 +05:30
meili-bors[bot]	76c05d1b20	Merge #4053 4053: Fix the stats of the documents deletion by filter r=Kerollmops a=irevoire # Pull Request The issue was that the operation « DocumentDeletionByFilter » was not declared as an index operation. That means the index stats were not reprocessed after the application of the operation. ## Related issue Fixes #4018 ## What does this PR do? - Move the `DocumentDeletionByFilter` internal operation into the category of the `IndexOperation`. This means that the stats will automatically be re-processed after a batch is processed. - Update a test to ensure that the stats are valid after each operation ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Tamo <tamo@meilisearch.com>	2023-09-11 15:53:26 +00:00
meili-bors[bot]	ef31ab52a4	Merge #4051 4051: Implement the snapshots on demand r=Kerollmops a=irevoire # Pull Request Private link: [PRD available here](https://www.notion.so/meilisearch/On-demand-snapshots-5676e542b905459d96eec228da133b00#847ff0cafeb64fe09e8ee7150852b474) Specification here: https://github.com/meilisearch/specifications/pull/258 ## Prototype A prototype is available under the name: `prototype-snapshot-on-demand-0`. ## Related issue Fixes #4052 ## What does this PR do? - Introduce a new route, `POST /snapshots` to create snapshots on demand - Introduce a new api-key action `snapshot.create` - Introduce a new analytic `Snapshot Created` sent every time a snapshot is created. ## Notes for the team I made a prototype so users can test the feature before the v1.5 comes out. But we can merge the PR as-is. Co-authored-by: Tamo <tamo@meilisearch.com>	2023-09-11 15:16:08 +00:00
Tamo	34fac115d5	fix clippy	2023-09-11 17:15:57 +02:00
Tamo	791c5cd874	makes clippy happy	2023-09-11 17:02:01 +02:00
Tamo	5bea1092fb	fix the flaky test	2023-09-11 16:56:26 +02:00
Tamo	056b2c387d	refactor the tests suite slightly	2023-09-11 16:56:26 +02:00
meili-bors[bot]	a09686fcbd	Merge #3997 3997: Refactor empty arrays/objects should return empty instead of null r=Kerollmops a=dogukanakkaya # Pull Request ## What does this PR do? At the moment if we select empty objects and array of object properties with dot notations like: ```json { "array": [], "object": {} } ``` ```rs GetDocumentOptions { fields: Some(vec!["array.name", "object.name"]) } ``` returns null if the array/object has no property yet. I am not sure if this is expected or it's the correct behaviour but I add my document with a property that is assigned to an empty array/object, later on when I select it, returns null which is kinda weird and unexpected in my opinion. This PR fixes that issue by returning an empty vector if the array is empty or an empty map if object is empty. This is not added for `permissive-json-pointer/src/lib.rs:224` because `create_array` loops over each item. Selecting a single property that is an object, in an array of objects would result other objects to be empty maps instead of none. ```json "doggos": [ { "jean": { "race": { "name": "bernese mountain", } } }, { "marc": { "age": 4, "race": { "name": "golden retriever", } } } ] ``` ```rs GetDocumentOptions { fields: Some(vec!["doggos.jean"]) } ``` Would result in `jean` object and an extra empty object for `marc`. ## PR checklist Please check if your PR fulfills the following requirements: - [ ] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: dogukanakkaya <doguakkaya27@hotmail.com>	2023-09-11 13:46:02 +00:00
meili-bors[bot]	b4c44603db	Merge #4009 4009: Bump rustls-webpki from 0.100.1 to 0.100.2 r=Kerollmops a=dependabot[bot] Bumps [rustls-webpki](https://github.com/rustls/webpki) from 0.100.1 to 0.100.2. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/rustls/webpki/releases">rustls-webpki's releases</a>.</em></p> <blockquote> <h2>v/0.100.2</h2> <h2>Release notes</h2> <ul> <li>certificate path building and verification is now capped at 100 signature validation operations to avoid the risk of CPU usage denial-of-service attack when validating crafted certificate chains producing quadratic runtime. This risk affected both clients, as well as servers that verified client certificates.</li> </ul> <h2>What's Changed</h2> <ul> <li>v0.100.2 prep by <a href="https://github.com/cpu"><code>`@cpu</code></a>` in <a href="https://redirect.github.com/rustls/webpki/pull/154">rustls/webpki#154</a></li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/rustls/webpki/compare/v/0.100.1...v/0.100.2">https://github.com/rustls/webpki/compare/v/0.100.1...v/0.100.2</a></p> </blockquote> </details> <details> <summary>Commits</summary> <ul> <li><a href="`c8b821450b`"><code>c8b8214</code></a> Bump MSRV to 1.60</li> <li><a href="`855752292e`"><code>8557522</code></a> Avoid testing MSRV of dev-dependencies</li> <li><a href="`73a7f0c7d7`"><code>73a7f0c</code></a> Cargo: version 0.100.1 -> 0.100.2</li> <li><a href="`4ea052366f`"><code>4ea0523</code></a> verify_cert: enforce maximum number of signatures.</li> <li>See full diff in <a href="https://github.com/rustls/webpki/compare/v/0.100.1...v/0.100.2">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=rustls-webpki&package-manager=cargo&previous-version=0.100.1&new-version=0.100.2)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) Dependabot will resolve any conflicts with this PR as long as you don't alter it yourself. You can also trigger a rebase manually by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` show <dependency name> ignore conditions` will show all of the ignore conditions of the specified dependency - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) You can disable automated security fix PRs for this repo from the [Security Alerts page](https://github.com/meilisearch/meilisearch/network/alerts). </details> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2023-09-11 13:11:07 +00:00
dogukanakkaya	393be40179	Refactor empty arrays/objects should return empty instead of null	2023-09-11 15:56:15 +03:00
Tamo	2c1d60f79b	get rid of a warning	2023-09-11 14:40:22 +02:00
meili-bors[bot]	487d493f49	Merge #4043 4043: Bring back hotfixes from v1.3.3 into v1.4.0 r=Kerollmops a=curquiza Co-authored-by: curquiza <curquiza@users.noreply.github.com> Co-authored-by: meili-bors[bot] <89034592+meili-bors[bot]@users.noreply.github.com> Co-authored-by: Kerollmops <clement@meilisearch.com> Co-authored-by: curquiza <clementine@meilisearch.com>	2023-09-11 12:27:34 +00:00
Tamo	08af69a33b	improve a test to understand what's going on with the ci	2023-09-11 14:23:57 +02:00
Tamo	9258e5b5bf	Fix the stats of the documents deletion by filter The issue was that the operation « DocumentDeletionByFilter » was not declared as an index operation. That means the indexes stats were not reprocessed after the application of the operation.	2023-09-11 14:04:10 +02:00
Tamo	ddd34a488a	update the api-key tests	2023-09-11 13:52:07 +02:00
meili-bors[bot]	526c2b3602	Merge #4050 4050: Bump webpki from 0.22.0 to 0.22.1 r=Kerollmops a=dependabot[bot] Bumps [webpki](https://github.com/briansmith/webpki) from 0.22.0 to 0.22.1. <details> <summary>Commits</summary> <ul> <li>See full diff in <a href="https://github.com/briansmith/webpki/commits">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=webpki&package-manager=cargo&previous-version=0.22.0&new-version=0.22.1)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) Dependabot will resolve any conflicts with this PR as long as you don't alter it yourself. You can also trigger a rebase manually by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` show <dependency name> ignore conditions` will show all of the ignore conditions of the specified dependency - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) You can disable automated security fix PRs for this repo from the [Security Alerts page](https://github.com/meilisearch/meilisearch/network/alerts). </details> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2023-09-11 11:46:22 +00:00
Tamo	e8c9367686	implement the snapshots on demand	2023-09-11 12:35:57 +02:00
dependabot[bot]	9636c5f558	Bump webpki from 0.22.0 to 0.22.1 Bumps [webpki](https://github.com/briansmith/webpki) from 0.22.0 to 0.22.1. - [Commits](https://github.com/briansmith/webpki/commits) --- updated-dependencies: - dependency-name: webpki dependency-type: indirect ... Signed-off-by: dependabot[bot] <support@github.com>	2023-09-11 10:32:34 +00:00
Harsh Upadhyay	b310830b5d	Improve test-suite.yml for CI failing when disabling tokenization (#4005 ) * [Update] test-suite.yml Added New run command for cargo tree without default features using if-then block * [Updated] test-disabled-tokenization in test-suite.yml * [Updated] test-suite.yml * Update .github/workflows/test-suite.yml --------- Co-authored-by: Clémentine U. - curqui <clementine@meilisearch.com>	2023-09-11 12:30:53 +02:00
meili-bors[bot]	462b4654c4	Merge #4028 4028: Fix highlighting bug when searching for a phrase with cropping r=ManyTheFish a=vivek-26 # Pull Request ## Related issue Fixes #3975 ## What does this PR do? This PR - - Fixes the bug where searching only for a phrase (containing multiple words) along with cropping, highlighted only the first word of the phrase. - Adds unit test case for the above mentioned scenario. ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Vivek Kumar <vivek.26@outlook.com>	2023-09-11 07:58:41 +00:00
Vivek Kumar	abfa7ded25	use a new temp index in the test	2023-09-08 12:32:47 +05:30
Vivek Kumar	f2837aaec2	add another test case	2023-09-08 11:39:54 +05:30
Vivek Kumar	11df155598	fix highlighting bug when searching for a phrase with cropping	2023-09-08 11:39:52 +05:30
curquiza	651657c03e	Fix git conflicts	2023-09-07 16:48:13 +02:00
meili-bors[bot]	b9ad59c969	Merge #4041 4041: Register the swap indexe task in a spawn blocking to be sure to never… r=ManyTheFish a=irevoire # Pull Request ## Related issue Fixes https://github.com/meilisearch/meilisearch/issues/4040 ## What does this PR do? - Register the swap indexes task in a spawn blocking task Co-authored-by: Tamo <tamo@meilisearch.com>	2023-09-07 10:22:01 +00:00
Tamo	66aa682e23	Register the swap indexe task in a spawn blocking to be sure to never block the main thread	2023-09-07 11:37:02 +02:00
meili-bors[bot]	256cf33bca	Merge #4039 4039: Fix multiple vectors dimensions r=ManyTheFish a=Kerollmops This PR fixes #4035, making providing multiple vectors in documents possible. This is fixed by extracting the vectors from the non-flattened version of the documents. Co-authored-by: Kerollmops <clement@meilisearch.com>	2023-09-07 09:25:58 +00:00
meili-bors[bot]	9945cbf9db	Merge #4038 4038: Fix filter escaping issues r=ManyTheFish a=Kerollmops This PR fixes #4034 by always escaping the sequences. Users must always put quotes (simple or double) to escape the filter values. Co-authored-by: Kerollmops <clement@meilisearch.com>	2023-09-06 12:29:29 +00:00
Kerollmops	03d0f628bd	Use the unescaper crate to unescape any char sequence	2023-09-06 13:59:45 +02:00
Kerollmops	ea78060916	Fix tests that were supposed to escape characters	2023-09-06 13:59:45 +02:00
Kerollmops	b42d48187a	Add a test case scenario	2023-09-06 13:59:44 +02:00
Kerollmops	679c0b0f97	Extract the vectors from the non-flattened version of the documents	2023-09-06 12:26:00 +02:00
Kerollmops	e02d0064bd	Add a test case scenario	2023-09-06 12:26:00 +02:00
meili-bors[bot]	7ef3572f11	Merge #4037 4037: Update version for the next release (v1.3.3) in Cargo.toml r=curquiza a=meili-bot ⚠️ This PR is automatically generated. Check the new version is the expected one and Cargo.lock has been updated before merging. Co-authored-by: curquiza <curquiza@users.noreply.github.com>	2023-09-06 09:50:58 +00:00
curquiza	93285041a9	Update version for the next release (v1.3.3) in Cargo.toml	2023-09-06 09:23:20 +00:00
meili-bors[bot]	dc3d9c90d9	Merge #3994 3994: Fix synonyms with separators r=Kerollmops a=ManyTheFish # Pull Request ## Related issue Fixes #3977 ## Available prototype ``` $ docker pull getmeili/meilisearch:prototype-fix-synonyms-with-separators-0 ``` ## What does this PR do? - add a new test - filter the empty synonyms after normalization Co-authored-by: ManyTheFish <many@meilisearch.com>	2023-09-05 14:42:46 +00:00
meili-bors[bot]	287cf25d39	Merge #4033 4033: Fix thai synonyms r=Kerollmops a=Kerollmops Fixes #4031 Co-authored-by: Kerollmops <clement@meilisearch.com> Co-authored-by: ManyTheFish <many@meilisearch.com>	2023-09-05 13:54:33 +00:00
ManyTheFish	66aa6d5871	Ignore tokens with empty normalized value during indexing process	2023-09-05 15:44:14 +02:00
Kerollmops	8ac5b765bc	Fix synonyms normalization	2023-09-04 16:12:48 +02:00
meili-bors[bot]	cea93e9a37	Merge #4016 4016: Define the full Homebrew formula path r=curquiza a=Kerollmops This PR fixes #4015 by defining the full Homebrew formula path. Co-authored-by: Clément Renault <clement@meilisearch.com>	2023-09-04 13:10:28 +00:00
Kerollmops	085aad0a94	Add a test	2023-09-04 14:39:33 +02:00
meili-bors[bot]	e9b62aacb3	Merge #4025 4025: Bump Swatinem/rust-cache from 2.5.1 to 2.6.2 r=curquiza a=dependabot[bot] Bumps [Swatinem/rust-cache](https://github.com/swatinem/rust-cache) from 2.5.1 to 2.6.2. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/swatinem/rust-cache/releases">Swatinem/rust-cache's releases</a>.</em></p> <blockquote> <h2>v2.6.2</h2> <h2>What's Changed</h2> <ul> <li>dep: Use <code>smol-toml</code> instead of <code>toml</code> by <a href="https://github.com/NobodyXu"><code>`@NobodyXu</code></a>` in <a href="https://redirect.github.com/Swatinem/rust-cache/pull/164">Swatinem/rust-cache#164</a></li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/Swatinem/rust-cache/compare/v2...v2.6.2">https://github.com/Swatinem/rust-cache/compare/v2...v2.6.2</a></p> <h2>v2.6.1</h2> <ul> <li>Fix hash contributions of <code>Cargo.lock</code>/<code>Cargo.toml</code> files.</li> </ul> <h2>v2.6.0</h2> <h2>What's Changed</h2> <ul> <li>Add "buildjet" as a second <code>cache-provider</code> backend <a href="https://github.com/joroshiba"><code>`@joroshiba</code></a>` in <a href="https://redirect.github.com/Swatinem/rust-cache/pull/154">Swatinem/rust-cache#154</a></li> <li>Clean up sparse registry index.</li> <li>Do not clean up src of <code>-sys</code> crates.</li> <li>Remove <code>.cargo/credentials.toml</code> before saving.</li> </ul> <h2>New Contributors</h2> <ul> <li><a href="https://github.com/joroshiba"><code>`@joroshiba</code></a>` made their first contribution in <a href="https://redirect.github.com/Swatinem/rust-cache/pull/154">Swatinem/rust-cache#154</a></li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/Swatinem/rust-cache/compare/v2.5.1...v2.6.0">https://github.com/Swatinem/rust-cache/compare/v2.5.1...v2.6.0</a></p> </blockquote> </details> <details> <summary>Changelog</summary> <p><em>Sourced from <a href="https://github.com/Swatinem/rust-cache/blob/master/CHANGELOG.md">Swatinem/rust-cache's changelog</a>.</em></p> <blockquote> <h2>2.6.2</h2> <ul> <li>Fix <code>toml</code> parsing.</li> </ul> <h2>2.6.1</h2> <ul> <li>Fix hash contributions of <code>Cargo.lock</code>/<code>Cargo.toml</code> files.</li> </ul> <h2>2.6.0</h2> <ul> <li>Add "buildjet" as a second <code>cache-provider</code> backend.</li> <li>Clean up sparse registry index.</li> <li>Do not clean up src of <code>-sys</code> crates.</li> <li>Remove <code>.cargo/credentials.toml</code> before saving.</li> </ul> </blockquote> </details> <details> <summary>Commits</summary> <ul> <li><a href="`e207df5d26`"><code>e207df5</code></a> 2.6.2</li> <li><a href="`decb69d790`"><code>decb69d</code></a> Update dependencies and add changelog</li> <li><a href="`ab6b2769d1`"><code>ab6b276</code></a> dep: Use <code>smol-toml</code> instead of <code>toml</code> (<a href="https://redirect.github.com/swatinem/rust-cache/issues/164">#164</a>)</li> <li><a href="`578b235f6e`"><code>578b235</code></a> 2.6.1</li> <li><a href="`5113490c3f`"><code>5113490</code></a> prepare 2.6.1</li> <li><a href="`c0e052c18c`"><code>c0e052c</code></a> Fix hashing of parsed <code>Cargo.toml</code> (<a href="https://redirect.github.com/swatinem/rust-cache/issues/160">#160</a>)</li> <li><a href="`4e0f4b19dd`"><code>4e0f4b1</code></a> Fix typo in hashing parsed <code>Cargo.lock</code> (<a href="https://redirect.github.com/swatinem/rust-cache/issues/159">#159</a>)</li> <li><a href="`b919e1427f`"><code>b919e14</code></a> feat: Add logging to <code>Cargo.lock</code>/<code>Cargo.toml</code> hashing (<a href="https://redirect.github.com/swatinem/rust-cache/issues/156">#156</a>)</li> <li><a href="`b8a6852b4f`"><code>b8a6852</code></a> 2.6.0</li> <li><a href="`80c47cc945`"><code>80c47cc</code></a> Clean up <code>credentials.toml</code></li> <li>Additional commits viewable in <a href="https://github.com/swatinem/rust-cache/compare/v2.5.1...v2.6.2">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=Swatinem/rust-cache&package-manager=github_actions&previous-version=2.5.1&new-version=2.6.2)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) You can trigger a rebase of this PR by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` show <dependency name> ignore conditions` will show all of the ignore conditions of the specified dependency - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) </details> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2023-09-04 12:30:53 +00:00
dependabot[bot]	456960d2c7	Bump Swatinem/rust-cache from 2.5.1 to 2.6.2 Bumps [Swatinem/rust-cache](https://github.com/swatinem/rust-cache) from 2.5.1 to 2.6.2. - [Release notes](https://github.com/swatinem/rust-cache/releases) - [Changelog](https://github.com/Swatinem/rust-cache/blob/master/CHANGELOG.md) - [Commits](https://github.com/swatinem/rust-cache/compare/v2.5.1...v2.6.2) --- updated-dependencies: - dependency-name: Swatinem/rust-cache dependency-type: direct:production update-type: version-update:semver-minor ... Signed-off-by: dependabot[bot] <support@github.com>	2023-09-01 17:17:39 +00:00
meili-bors[bot]	3dda176723	Merge #4020 4020: Update version for the next release (v1.4.0) in Cargo.toml r=Kerollmops a=meili-bot ⚠️ This PR is automatically generated. Check the new version is the expected one and Cargo.lock has been updated before merging. Co-authored-by: Kerollmops <Kerollmops@users.noreply.github.com> Co-authored-by: Clément Renault <clement@meilisearch.com>	2023-08-28 13:51:23 +00:00
Clément Renault	af0f6f0bf0	Merge branch 'main' into update-version-v1.4.0	2023-08-28 15:08:59 +02:00
meili-bors[bot]	ccf3ba3f32	Merge #4019 4019: Bringing back changes from `v1.3.2` onto `main` r=irevoire a=Kerollmops Co-authored-by: Kerollmops <clement@meilisearch.com> Co-authored-by: meili-bors[bot] <89034592+meili-bors[bot]@users.noreply.github.com> Co-authored-by: irevoire <irevoire@users.noreply.github.com> Co-authored-by: Clément Renault <clement@meilisearch.com>	2023-08-28 12:14:11 +00:00
Kerollmops	65528a3e06	Update version for the next release (v1.4.0) in Cargo.toml	2023-08-28 11:52:28 +00:00
Clément Renault	6db80b0836	Define the full Homebrew formula path	2023-08-24 11:24:47 +02:00
meili-bors[bot]	cdb4b3e024	Merge #4013 4013: Fix the ranking rule by temporarily disabling an assert in the bucket sort algorithm r=Kerollmops a=Kerollmops This PR temporarily disables an assertion, making the search crash. [I created a tracking issue](https://github.com/meilisearch/meilisearch/issues/4012) to find a better way to fix this. It no longer reverts `a20e4d447c`, which seemed to generate unreachable graphs and make the bucket sort ranking algorithm panic because of entering an unreachable state. We discussed that below in the comments. Temporary fixes #4002, fixes #4006, and fixes #3995. --- It took me approximately 2 days to find the first bad commit just because I'm bad in `git bisect` x `bash`, i.e. [I misused `%1` with `$!` to kill the most recently backgrounded job](https://unix.stackexchange.com/a/340084/212574)... <details> <summary>Here is the script I used to find the invalid commit</summary> ```bash #!/usr/bin/env bash set -x # remove the data rm -rf data.ms # build meilisearch cargo build --release # ignore this commit if it doesn't compile if [[ $? != 0 ]]; then exit 125 fi # index the dump and start from it ./target/release/meilisearch \ --http-addr 'localhost:7705' \ --import-dump $HOME/Downloads/modified-20230822-083016113.dump & # wait 10 sec while it indexes the docs sleep 5 # check if the server crashes on requests echo '{ "q": "rtx 305", "attributesToHighlight": [ "*" ], "highlightPreTag": "<ais-highlight-0000000000>", "highlightPostTag": "</ais-highlight-0000000000>", "limit": 21, "offset": 0 }' \| xh 'localhost:7705/indexes/arvutitark_local_orderables/search' last_exit_code=$? # Now kill Meilisearch kill $! # Clean the potential Cargo.lock git checkout . exit $last_exit_code ``` </details> Co-authored-by: Kerollmops <clement@meilisearch.com> Co-authored-by: Clément Renault <clement@meilisearch.com>	2023-08-23 15:30:56 +00:00
Clément Renault	8c0ebd1331	Update milli/src/search/new/bucket_sort.rs Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-08-23 16:40:39 +02:00
Kerollmops	5130e06b41	Temporarily disable an assert in the ranking rules	2023-08-23 16:11:54 +02:00
Clément Renault	08e27ef73f	Merge pull request #4008 from meilisearch/fix-highlighting-panic Bump charabia to 0.8.3	2023-08-23 11:56:45 +02:00
meili-bors[bot]	914b125c5f	Merge #3945 3945: Do not leak field information on error r=Kerollmops a=vivek-26 # Pull Request ## Related issue Fixes #3865 ## What does this PR do? This PR ensures that `InvalidSortableAttribute`and `InvalidFacetSearchFacetName` errors do not leak field information i.e. fields which are not part of `displayedAttributes` in the settings are hidden from the error message. ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Vivek Kumar <vivek.26@outlook.com>	2023-08-22 18:55:27 +00:00
dependabot[bot]	e59d7f238c	Bump rustls-webpki from 0.100.1 to 0.100.2 Bumps [rustls-webpki](https://github.com/rustls/webpki) from 0.100.1 to 0.100.2. - [Release notes](https://github.com/rustls/webpki/releases) - [Commits](https://github.com/rustls/webpki/compare/v/0.100.1...v/0.100.2) --- updated-dependencies: - dependency-name: rustls-webpki dependency-type: indirect ... Signed-off-by: dependabot[bot] <support@github.com>	2023-08-22 18:10:53 +00:00
Kerollmops	717b069907	Bump charabia to 0.8.3	2023-08-22 16:25:00 +02:00
meili-bors[bot]	7ea154673a	Merge #4000 4000: Update version for the next release (v1.3.2) in Cargo.toml r=irevoire a=meili-bot ⚠️ This PR is automatically generated. Check the new version is the expected one and Cargo.lock has been updated before merging. Co-authored-by: irevoire <irevoire@users.noreply.github.com>	2023-08-16 10:41:33 +00:00
irevoire	b947f3bb9d	Update version for the next release (v1.3.2) in Cargo.toml	2023-08-16 08:20:36 +00:00
meili-bors[bot]	4c35817c5f	Merge #3998 3998: Accept the `null` JSON value as a value of the `_vectors` field r=irevoire a=Kerollmops This PR fixes #3979 by accepting `null` JSON values in the `_vectors` fields provided by the user. Can the reviewer please verify that I am merging in the right branch? I think we must create a new _release-v1.3.2_. Co-authored-by: Kerollmops <clement@meilisearch.com>	2023-08-16 08:12:24 +00:00
Kerollmops	c53841e166	Accept the null JSON value as the value of _vectors	2023-08-14 16:03:55 +02:00
meili-bors[bot]	fd81945597	Merge #3987 3987: Update dependencies for v1.4 r=curquiza a=ManyTheFish # Pull Request ## Related issue Fixes #3870 ## What does this PR do? - [Update dependencies](`d7ff5368b4`) - [upgrade itertools = "0.10.5"](`d0582d01f4`) - [upgrade sysinfo = "0.29.7"](`507c661352`) - [upgrade memmap2 = "0.7.1"](`489e0d5cd0`) - [upgrade rstar = "0.11.0"](`3d9d08e3b2`) - [upgrade fastrand = "2.0.0"](`1af7083c48`) - [upgrade deserr = "0.6.0"](`7fe77045af`) - [upgrade indexmap = "2.0.0"](`95e4960b0c`) - [update rust toolchain = "1.71.1"](`937b7b5da5`) ## Remaining un-upgraded dependencies - vergen 7.5.1 --> 8.2.4: I wasn't able to quickly understand the changes in the lib API to upgrade the dependency - rustls 0.20.8 --> 0.21.6: Meilisearch doesn't have any direct dependency on it Co-authored-by: ManyTheFish <many@meilisearch.com>	2023-08-10 16:46:17 +00:00
ManyTheFish	794e491152	update rust toolchain	2023-08-10 18:09:02 +02:00
ManyTheFish	cab27c2ab4	upgrade indexmap = "2.0.0"	2023-08-10 18:09:02 +02:00
ManyTheFish	624fa9052f	upgrade deserr = "0.6.0"	2023-08-10 18:09:02 +02:00
ManyTheFish	359ede4862	upgrade fastrand = "2.0.0"	2023-08-10 18:09:02 +02:00
ManyTheFish	60c11dbdbd	upgrade rstar - "0.11.0"	2023-08-10 18:09:02 +02:00
ManyTheFish	dacee40ebc	upgrade memmap2 = "0.7.1"	2023-08-10 18:09:02 +02:00
ManyTheFish	6089083a8e	upgrade sysinfo = "0.29.7"	2023-08-10 18:09:02 +02:00
ManyTheFish	cc2c19d4c3	upgrade itertools = "0.10.5"	2023-08-10 18:09:02 +02:00
ManyTheFish	a5c56fac8a	Update dependencies	2023-08-10 18:09:02 +02:00
meili-bors[bot]	e4e49e63d0	Merge #3993 3993: Bringing back changes from v1.3.1 to `main` r=irevoire a=curquiza Co-authored-by: irevoire <irevoire@users.noreply.github.com> Co-authored-by: meili-bors[bot] <89034592+meili-bors[bot]@users.noreply.github.com> Co-authored-by: Tamo <tamo@meilisearch.com> Co-authored-by: ManyTheFish <many@meilisearch.com>	2023-08-10 14:30:02 +00:00
meili-bors[bot]	00bd7bd19a	Merge #3990 3990: Removed unnecessary borrow call that failed nightly tests r=irevoire a=JannisK89 # Pull Request ## Related issue Fixes #3988 ## What does this PR do? - Removes unnecessary borrow call that was causing warnings when running tests on nightly. ## PR checklist Please check if your PR fulfills the following requirements: - [ x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [ x] Have you read the contributing guidelines? - [ x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Please let me know if there is anything else I can do to improve this PR. Thank you. Co-authored-by: JannisK89 <jannis.karanikis@gmail.com>	2023-08-10 11:42:19 +00:00
meili-bors[bot]	ef3d098b4d	Merge #3976 3976: Fix the get stats method r=ManyTheFish a=irevoire # Pull Request - The get stats method of the index-scheduler was not using at all the processing tasks. That was returning a wrong number of enqueued tasks and 0 processing tasks. - Added a test - Currently this method was ONLY used to compute the `meilisearch_nb_tasks` field of the experimental feature metrics. ## Related issue Fixes https://github.com/meilisearch/meilisearch/issues/3972 Co-authored-by: Tamo <tamo@meilisearch.com>	2023-08-10 10:55:50 +00:00
meili-bors[bot]	8084cf29f3	Merge #3946 3946: Settings customizing tokenization r=irevoire a=ManyTheFish # Pull Request This pull Request allows the User to customize Meilisearch Tokenization by providing specialized settings. ## Small documentation All the new settings can be set and reset like the other index settings by calling the route `/indexes/:name/settings` ### `nonSeparatorTokens` The Meilisearch word segmentation uses a default list of separators to segment words, however, for specific use cases some of the default separators shouldn't be considered separators, the `nonSeparatorTokens` setting allows to remove of some tokens from the default list of separators. *Request payload `PUT`- `/indexes/articles/settings/non-separator-tokens`* ```json ["`@",` "#", "&"] ``` ### `separatorTokens` Some use cases need to define additional separators, some are related to a specific way of parsing technical documents some others are related to encodings in documents, the `separatorTokens` setting allows adding some tokens to the list of separators. *Request payload `PUT`- `/indexes/articles/settings/separator-tokens`* ```json ["§", "&sep"] ``` ### `dictionary` The Meilisearch word segmentation relies on separators and language-based word-dictionaries to segment words, however, this segmentation is inaccurate on technical or use-case specific vocabulary (like `G/Box` to say `Gear Box`), or on proper nouns (like `J. R. R.` when parsing `J. R. R. Tolkien`), the `dictionary` setting allows defining a list of words that would be segmented as described in the list. *Request payload `PUT`- `/indexes/articles/settings/dictionary`* ```json ["J. R. R.", "J.R.R."] ``` these last feature synergies well with the `stopWords` setting or the `synonyms` setting allowing to segment words and correctly retrieve the synonyms: *Request payload `PATCH`- `/indexes/articles/settings`* ```json { "dictionary": ["J. R. R.", "J.R.R."], "synonyms": { "J.R.R.": ["jrr", "J. R. R."], "J. R. R.": ["jrr", "J.R.R."], "jrr": ["J.R.R.", "J. R. R."], } } ``` ### Related specifications: - https://github.com/meilisearch/specifications/pull/255 - https://github.com/meilisearch/specifications/pull/254 ### Try it with Docker ```bash $ docker pull getmeili/meilisearch:prototype-tokenizer-customization-3 ``` ## Related issue Fixes #3610 Fixes #3917 Fixes https://github.com/meilisearch/product/discussions/468 Fixes https://github.com/meilisearch/product/discussions/160 Fixes https://github.com/meilisearch/product/discussions/260 Fixes https://github.com/meilisearch/product/discussions/381 Fixes https://github.com/meilisearch/product/discussions/131 Related to https://github.com/meilisearch/meilisearch/issues/2879 Fixes #2760 ## What does this PR do? - Add a setting `nonSeparatorTokens` allowing to remove a token from the default separator tokens - Add a setting `separatorTokens` allowing to add a token in the separator tokens - Add a setting `dictionary` allowing to override the segmentation on specific words - add new error code `invalid_settings_non_separator_tokens` (invalid_request) - add new error code `invalid_settings_separator_tokens` (invalid_request) - add new error code `invalid_settings_dictionary` (invalid_request) Co-authored-by: ManyTheFish <many@meilisearch.com> Co-authored-by: Many the fish <many@meilisearch.com>	2023-08-10 10:01:18 +00:00
ManyTheFish	5a7c1bde84	Fix clippy	2023-08-10 11:27:56 +02:00
ManyTheFish	6b2d671be7	Fix PR comments	2023-08-10 10:44:07 +02:00
Many the fish	43c13faeda	Update milli/src/update/index_documents/extract/extract_docid_word_positions.rs Co-authored-by: Tamo <tamo@meilisearch.com>	2023-08-10 10:05:03 +02:00
meili-bors[bot]	29adfc2f68	Merge #3989 3989: Improve test suite CI for manual trigger events r=irevoire a=curquiza # Why? To be able to test https://github.com/meilisearch/meilisearch/issues/3988 before merging the PR solving it # How do we ensure this PR works? I triggered `workflow_dispatch` (i.e. manual trigger) on this branch, and we can see all the jobs have been triggered (even if some of them are failing -> it's another issue) https://github.com/meilisearch/meilisearch/actions/runs/5810609073 We can see the tests triggered by the PR are restricted as expected: https://github.com/meilisearch/meilisearch/actions/runs/5810605977 Co-authored-by: curquiza <clementine@meilisearch.com>	2023-08-10 07:55:48 +00:00
JannisK89	064ee95b1c	removed unnecessary borrow call	2023-08-10 08:41:25 +02:00
curquiza	604d533b31	Improve test suite CI for workflow_dispatch event	2023-08-09 16:47:28 +02:00
meili-bors[bot]	44c1900f36	Merge #3986 3986: Fix geo bounding box with strings r=ManyTheFish a=irevoire # Pull Request When sending a document with one geofield of type string (i.e.: `{ "_geo": { "lat": 12, "lng": "13" }}`), the geobounding box would exclude this document. This PR fixes this issue by automatically parsing the string value in case we're working on a geofield. ## Related issue Fixes https://github.com/meilisearch/meilisearch/issues/3973 ## What does this PR do? - Automatically parse the facet value iif we're working on a geofield. - Make insta works with snapshots in loops or closure executed multiple times. (you may need to update your cli if it panics after this PR: `cargo install cargo-insta`). - Add one integration test in milli and in meilisearch to ensure it works forever. - Add three snapshots for the dump that mysteriously disappeared I don't know how Co-authored-by: Tamo <tamo@meilisearch.com>	2023-08-09 07:58:15 +00:00
meili-bors[bot]	04671d0751	Merge #3981 3981: Truncate the normalized long facets used in the search for facet value r=irevoire a=ManyTheFish # Pull Request Truncate the normalized long facets used in the search for facet value ## targeted release v1.3.1 ## Related issue Fixes #3978 Co-authored-by: ManyTheFish <many@meilisearch.com>	2023-08-08 15:07:07 +00:00
Tamo	4f4c669d50	add back some dump snapshots that disappeared. it's completely unrelated to this PR	2023-08-08 16:58:14 +02:00
ManyTheFish	8dc5acf998	Try fix	2023-08-08 16:52:36 +02:00
ManyTheFish	fc2590fc9d	Add a test	2023-08-08 16:43:08 +02:00
ManyTheFish	35758db9ec	Truncate the the normalized long facets used in search for facet value	2023-08-08 16:38:30 +02:00
Tamo	4988199bb9	ensure the geoboundingbox works with strings and int geofields in milli and meilisearch	2023-08-08 16:29:25 +02:00
Tamo	83991ee770	enable the multi-snapshot attribute in insta. This will let us use insta in loops	2023-08-08 16:28:38 +02:00
Tamo	9d061cec26	automatically parse the filterable attribute to float if it's a geo field	2023-08-08 16:28:07 +02:00
ManyTheFish	4a21fecf67	Merge branch 'main' into settings-customizing-tokenization	2023-08-08 16:08:16 +02:00
ManyTheFish	ae8e69c030	Add API route for the new settings	2023-08-08 16:03:16 +02:00
Tamo	fe819a9d80	fix the get stats method It was not taking into account the processing tasks at all	2023-08-08 13:21:15 +02:00
meili-bors[bot]	e338ceb97f	Merge #3982 3982: Update version for the next release (v1.3.1) in Cargo.toml r=irevoire a=meili-bot ⚠️ This PR is automatically generated. Check the new version is the expected one and Cargo.lock has been updated before merging. Co-authored-by: irevoire <irevoire@users.noreply.github.com>	2023-08-08 10:30:56 +00:00
irevoire	75c87d5391	Update version for the next release (v1.3.1) in Cargo.toml	2023-08-08 10:30:06 +00:00
Vivek Kumar	dd57873f8e	hide fields not in the displayedAttributes list from errors	2023-08-05 16:03:10 +05:30
meili-bors[bot]	3dda93d50f	Merge #3968 3968: Bump svenstaro/upload-release-action from 2.6.1 to 2.7.0 r=curquiza a=dependabot[bot] Bumps [svenstaro/upload-release-action](https://github.com/svenstaro/upload-release-action) from 2.6.1 to 2.7.0. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/svenstaro/upload-release-action/releases">svenstaro/upload-release-action's releases</a>.</em></p> <blockquote> <h2>2.7.0</h2> <ul> <li>Allow setting an explicit target_commitish <a href="https://redirect.github.com/svenstaro/upload-release-action/pull/46">#46</a> (thanks <a href="https://github.com/Spikatrix"><code>`@Spikatrix</code></a>)</li>` </ul> </blockquote> </details> <details> <summary>Changelog</summary> <p><em>Sourced from <a href="https://github.com/svenstaro/upload-release-action/blob/master/CHANGELOG.md">svenstaro/upload-release-action's changelog</a>.</em></p> <blockquote> <h2>[2.7.0] - 2023-07-28</h2> <ul> <li>Allow setting an explicit target_commitish <a href="https://redirect.github.com/svenstaro/upload-release-action/pull/46">#46</a> (thanks <a href="https://github.com/Spikatrix"><code>`@Spikatrix</code></a>)</li>` </ul> </blockquote> </details> <details> <summary>Commits</summary> <ul> <li><a href="`1beeb572c1`"><code>1beeb57</code></a> 2.7.0</li> <li><a href="`5206d34958`"><code>5206d34</code></a> Bump deps</li> <li><a href="`80d7a7e41c`"><code>80d7a7e</code></a> Merge pull request <a href="https://redirect.github.com/svenstaro/upload-release-action/issues/46">#46</a> from Spikatrix/master</li> <li><a href="`5eb2ffd70b`"><code>5eb2ffd</code></a> Merge pull request <a href="https://redirect.github.com/svenstaro/upload-release-action/issues/110">#110</a> from svenstaro/dependabot/npm_and_yarn/word-wrap-1.2.4</li> <li><a href="`07af2f374a`"><code>07af2f3</code></a> Bump word-wrap from 1.2.3 to 1.2.4</li> <li><a href="`5164410c7d`"><code>5164410</code></a> Push dist</li> <li><a href="`f47fb36ff1`"><code>f47fb36</code></a> Use the ref api to check if a tag exists</li> <li><a href="`212d4babf8`"><code>212d4ba</code></a> Rethrow getTag error if not 404</li> <li><a href="`7670b98fa0`"><code>7670b98</code></a> Push dist files</li> <li><a href="`ac438791c4`"><code>ac43879</code></a> Warn when target_commit is ignored</li> <li>Additional commits viewable in <a href="https://github.com/svenstaro/upload-release-action/compare/2.6.1...2.7.0">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=svenstaro/upload-release-action&package-manager=github_actions&previous-version=2.6.1&new-version=2.7.0)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) You can trigger a rebase of this PR by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) </details> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2023-08-02 09:55:39 +00:00
meili-bors[bot]	117146ec4e	Merge #3969 3969: Bump Swatinem/rust-cache from 2.5.0 to 2.5.1 r=curquiza a=dependabot[bot] Bumps [Swatinem/rust-cache](https://github.com/swatinem/rust-cache) from 2.5.0 to 2.5.1. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/swatinem/rust-cache/releases">Swatinem/rust-cache's releases</a>.</em></p> <blockquote> <h2>v2.5.1</h2> <ul> <li>Fix hash contribution of <code>Cargo.lock</code>.</li> </ul> </blockquote> </details> <details> <summary>Changelog</summary> <p><em>Sourced from <a href="https://github.com/Swatinem/rust-cache/blob/master/CHANGELOG.md">Swatinem/rust-cache's changelog</a>.</em></p> <blockquote> <h2>2.5.1</h2> <ul> <li>Fix hash contribution of <code>Cargo.lock</code>.</li> </ul> </blockquote> </details> <details> <summary>Commits</summary> <ul> <li><a href="`dd05243424`"><code>dd05243</code></a> 2.5.1</li> <li><a href="`65dbc54a5d`"><code>65dbc54</code></a> update changelog</li> <li><a href="`be7377e68e`"><code>be7377e</code></a> fix <code>src/config.ts</code>: Remove <code>sort_object</code> (<a href="https://redirect.github.com/swatinem/rust-cache/issues/152">#152</a>)</li> <li>See full diff in <a href="https://github.com/swatinem/rust-cache/compare/v2.5.0...v2.5.1">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=Swatinem/rust-cache&package-manager=github_actions&previous-version=2.5.0&new-version=2.5.1)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) You can trigger a rebase of this PR by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) </details> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2023-08-02 09:19:03 +00:00
dependabot[bot]	884b4d47b1	Bump Swatinem/rust-cache from 2.5.0 to 2.5.1 Bumps [Swatinem/rust-cache](https://github.com/swatinem/rust-cache) from 2.5.0 to 2.5.1. - [Release notes](https://github.com/swatinem/rust-cache/releases) - [Changelog](https://github.com/Swatinem/rust-cache/blob/master/CHANGELOG.md) - [Commits](https://github.com/swatinem/rust-cache/compare/v2.5.0...v2.5.1) --- updated-dependencies: - dependency-name: Swatinem/rust-cache dependency-type: direct:production update-type: version-update:semver-patch ... Signed-off-by: dependabot[bot] <support@github.com>	2023-08-01 17:22:43 +00:00
dependabot[bot]	023cb0c2de	Bump svenstaro/upload-release-action from 2.6.1 to 2.7.0 Bumps [svenstaro/upload-release-action](https://github.com/svenstaro/upload-release-action) from 2.6.1 to 2.7.0. - [Release notes](https://github.com/svenstaro/upload-release-action/releases) - [Changelog](https://github.com/svenstaro/upload-release-action/blob/master/CHANGELOG.md) - [Commits](https://github.com/svenstaro/upload-release-action/compare/2.6.1...2.7.0) --- updated-dependencies: - dependency-name: svenstaro/upload-release-action dependency-type: direct:production update-type: version-update:semver-minor ... Signed-off-by: dependabot[bot] <support@github.com>	2023-08-01 17:22:37 +00:00
meili-bors[bot]	f391039a6f	Merge #3967 3967: Bring back changes from `release-v1.3.0` into `main` r=ManyTheFish a=curquiza Using a temp branch because of git conflict Co-authored-by: Cong Chen <cong.chen@ocrlabs.com> Co-authored-by: ManyTheFish <many@meilisearch.com> Co-authored-by: Louis Dureuil <louis@meilisearch.com> Co-authored-by: meili-bors[bot] <89034592+meili-bors[bot]@users.noreply.github.com> Co-authored-by: Tamo <tamo@meilisearch.com> Co-authored-by: Kerollmops <clement@meilisearch.com> Co-authored-by: Clément Renault <clement@meilisearch.com>	2023-08-01 16:22:09 +00:00
curquiza	fcdd20b533	Fix README after git conflict	2023-08-01 16:06:33 +02:00
ManyTheFish	b45c36cd71	Merge branch 'main' into tmp-release-v1.3.0	2023-08-01 15:05:17 +02:00
meili-bors[bot]	151c31c18f	Merge #3963 3963: Fix the milli crate r=ManyTheFish a=irevoire Milli was using the serde feature of either without enabling it first; thus, it wasn't working. It was working in meilisearch, though, because `meilisearch-types` was using the feature which enables it globally for all the other crates. ## Related issue Fixes https://github.com/meilisearch/meilisearch/issues/3962 Co-authored-by: Tamo <tamo@meilisearch.com>	2023-07-31 09:32:08 +00:00
Tamo	a8ad0902d3	Fix the milli crate Milli was using the serde feature of either without enabling it first, thus it wasn't working	2023-07-31 11:08:27 +02:00
meili-bors[bot]	e917dbdebb	Merge #3957 3957: fix: upgrade mimalloc dependency to resolve FreeBSD build r=irevoire a=ThatOneCalculator # Pull Request ## Related issue Fixes #3806 ## What does this PR do? - Upgrades mimalloc to 0.1.37 - Fixes build on FreeBSD Ref: https://github.com/meilisearch/meilisearch/issues/3806#issuecomment-1653693468 Tested and working on FreeBSD 13.1-RELEASE-p5 ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: ThatOneCalculator <kainoa@t1c.dev>	2023-07-31 08:49:36 +00:00
ThatOneCalculator	ba919b6123	fix: ⬆️ up mimalloc	2023-07-28 20:35:47 -07:00
ManyTheFish	9d5e3457e5	Fix clippy	2023-07-27 14:21:19 +02:00
ManyTheFish	04694071fe	Fix the synonyms settings display	2023-07-27 14:12:23 +02:00
meili-bors[bot]	5b0157c6c6	Merge #3955 3955: Update mini-dashboard to version 0.2.11 r=curquiza a=bidoubiwa # Pull Request ## What does this PR do? - Updates the mini-dashboard to version [0.2.11](https://github.com/meilisearch/mini-dashboard/releases/tag/v0.2.11) ## PR checklist - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Charlotte Vermandel <charlottevermandel@gmail.com>	2023-07-27 11:59:55 +00:00
Charlotte Vermandel	3b9a87c790	Update mini-dashboard to version 0.2.11	2023-07-27 13:16:32 +02:00
meili-bors[bot]	3a3414270d	Merge #3952 3952: Use the new safe `read-txn-no-tls` heed feature r=ManyTheFish a=Kerollmops [We recently found out](https://github.com/meilisearch/heed/issues/191#issuecomment-1650280513) that the `read-sync-txn` heed feature was invalid and must be removed from this crate. We were declaring it in milli/meilisearch but, fortunately, not sharing the `RoTxn`s across threads 😮‍💨 [I recently introduced the `read-txn-no-tls` heed feature](https://github.com/meilisearch/heed/pull/194), which implements `RoTxn: Send` and allows multiple read transactions on a single thread (which we use). This PR removes the `sync-read-txn` heed feature from the _Cargo.toml_ file. I will fix this in heed v0.20.0 and will fill a RustSec advisory in the meantime. Co-authored-by: Clément Renault <clement@meilisearch.com>	2023-07-26 16:40:58 +00:00
meili-bors[bot]	d06e0905db	Merge #3953 3953: Update UTM campaign r=curquiza a=macraig # Pull Request ## What does this PR do? Redirect CTAs to Cloud landing page Co-authored-by: María <maria@Marias-MacBook-Pro.local>	2023-07-26 15:20:40 +00:00
meili-bors[bot]	939b2fc6fd	Merge #3949 3949: Fix score details casing r=Kerollmops a=ManyTheFish # Pull Request Fixes #3941 Co-authored-by: ManyTheFish <many@meilisearch.com>	2023-07-26 14:14:59 +00:00
María	fae61372be	Redirect CTAs to Cloud landing page	2023-07-26 15:54:43 +02:00
Clément Renault	d8b47b689e	Use the new read-txn-no-tls heed feature	2023-07-26 15:45:15 +02:00
ManyTheFish	b0c1a9504a	ensure the synonyms are updated when the tokenizer settings are changed	2023-07-26 09:33:42 +02:00
meili-bors[bot]	be72be7c0d	Merge #3942 3942: Normalize for the search the facets values r=ManyTheFish a=Kerollmops This PR improves and fixes the search for facet values feature. Searching for _bre_ wasn't returning facet values like _brévent_ or _brô_. The issue was related to the fact that facets are normalized but not in the same way as the `searchableAttributes` are. We decided to normalize them further and add another intermediate database where the key is the normalized facet value, and the value is a set of the non-normalized facets. We then use these non-normalized ones to get the correct counts by fetching the associated databases. ### What's missing in this PR? - [x] Apply the change to the whole set of `SearchForFacetValue::execute` conditions. - [x] Factorize the code that does an intermediate normalized value fetch in a function. - [x] Add or modify the search for facet value test. Co-authored-by: Clément Renault <clement@meilisearch.com> Co-authored-by: Kerollmops <clement@meilisearch.com>	2023-07-25 14:37:17 +00:00
ManyTheFish	88559a2d54	Fix score details casing	2023-07-25 15:49:33 +02:00
Clément Renault	59201a7852	Use snapshot instead of asserts Co-authored-by: Many the fish <many@meilisearch.com>	2023-07-25 15:34:05 +02:00
meili-bors[bot]	9e3e69373e	Merge #3948 3948: Fix hnsw internal panic by using another library r=ManyTheFish a=Kerollmops This pull request fixes #3923. The issue concerns the `hnsw` crate panicking due to a wrong call to the `[T]::copy_from_slice` function. I decided to switch the library to `instant-distance`, which is maintained [by someone of trust](https://lib.rs/~djc), who maintains a lot of very important crates. - [x] Make Clippy happy with the first commit. - [x] Reproduce the #3923 bug without this patch - [x] Check if the bug disappeared with this PR. - [x] Test with [the Algolia e-commerce dataset](https://www.notion.so/meilisearch/Algolia-Ecommerce-c5fa3b5f23a7485295df7e87306d5859). Co-authored-by: Kerollmops <clement@meilisearch.com>	2023-07-25 13:28:25 +00:00
ManyTheFish	d57026cd96	Support synonyms sinergies	2023-07-25 15:01:42 +02:00
Kerollmops	29ab54b259	Replace the hnsw crate by the instant-distance one	2023-07-25 12:37:35 +02:00
ManyTheFish	41c9e8856a	Fix test	2023-07-25 10:55:37 +02:00
Kerollmops	86d8bb3a3e	Make clippy happy (again)	2023-07-25 10:30:50 +02:00
ManyTheFish	d4ff59fcf5	Fix clippy	2023-07-24 18:42:26 +02:00
ManyTheFish	9c485f8563	Make the search and the indexing work	2023-07-24 18:35:20 +02:00
Kerollmops	0e2a5951b4	Add more advanced tests	2023-07-24 18:04:58 +02:00
Kerollmops	691a536893	Implement the facet search with the normalized index	2023-07-24 17:56:17 +02:00
ManyTheFish	d8d12d5979	Be able to set and reset settings	2023-07-24 17:00:18 +02:00
Clément Renault	df528b41d8	Normalize for the search the facets values	2023-07-20 17:57:07 +02:00
meili-bors[bot]	2452ec55b4	Merge #3940 3940: Update mini dashboard v0.2.9 r=gillian-meilisearch a=bidoubiwa # Pull Request ## What does this PR do? - Updates the mini-dashboard to version [0.2.9](https://github.com/meilisearch/mini-dashboard/releases/tag/v0.2.9) ## PR checklist - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Charlotte Vermandel <charlottevermandel@gmail.com>	2023-07-20 15:08:59 +00:00
Charlotte Vermandel	54ae1b5a67	Update mini-dashboard to version 0.2.9	2023-07-20 14:11:17 +02:00
ManyTheFish	0597a97c84	Update tests	2023-07-20 11:15:10 +02:00
meili-bors[bot]	3070a20580	Merge #3937 3937: Update Charabia to the last version r=Kerollmops a=ManyTheFish # Pull Request ## Related issue Fixes #3924 ## What does this PR do? - Update Charabia Co-authored-by: ManyTheFish <many@meilisearch.com>	2023-07-19 14:57:38 +00:00
ManyTheFish	0497f93494	Update Charabia to the last version	2023-07-19 15:19:32 +02:00
meili-bors[bot]	2dfbb6813a	Merge #3913 3913: Expose a Puffin server to profile the indexing process r=Kerollmops a=Kerollmops This PR exposes a puffin HTTP server to expose the internal timing it takes to index documents, delete documents, or update the settings of an index. <img width="1752" alt="Capture d’écran 2023-07-10 à 18 44 58" src="https://github.com/meilisearch/meilisearch/assets/3610253/a3c7a6bf-db5b-42f4-8be1-c4e31c869843"> ## To be done - [x] Move the puffin HTTP server under a feature flag. - [x] Use [the `puffin::set_scopes_on` function](https://docs.rs/puffin/latest/puffin/fn.set_scopes_on.html) to toggle it (by using the feature directly). When this function is called with `false`, [a call to `profile_scope!` talked 1-2ns](https://docs.rs/puffin/latest/puffin/fn.set_scopes_on.html). - [x] Create a _PROFILING.md_ file explaining how to use it. - [x] Explain that merging scopes on the interface is not always useful. - [x] Add more info on the number of batched tasks (using the `puffin::profile_scope!` macro data). - I added more info, but that's more continuous work when we consider we need more info here and there. - [x] Clean up some scopes, and don't touch too much code to inject puffin. - I am not sure that the _index_documents/mod.rs_ function is that complex with the addition of the scope. - [x] Think about what we consider frames. One indexation operation or the wall program. When must we stop the frame, then? - What we consider a frame is one single `IndexScheduler::tick` execution. - We can change that later. Co-authored-by: Kerollmops <clement@meilisearch.com> Co-authored-by: Clément Renault <clement@meilisearch.com>	2023-07-19 09:44:01 +00:00
Clément Renault	8f589a5cce	Introduce a PROFILING.md tutorial to profile Meilisearch	2023-07-18 17:38:13 +02:00
Clément Renault	0b8bbd8750	Toggle the puffin profiling with a feature flag	2023-07-18 17:38:13 +02:00
Kerollmops	eef95de30e	First iteration on exposing puffin profiling	2023-07-18 17:38:13 +02:00
meili-bors[bot]	13a13a4862	Merge #3932 3932: Add UTM tracking to README r=gillian-meilisearch a=Strift # Pull Request Hi `@macraig` `@curquiza` 👋 ## Related issue N/A ## What does this PR do? This PR adds UTM tracking to the links in the README. It add UTM params to: - links in the nav - links to where2watch - links in the Features section - Docs & Getting started links (cc `@guimachiavelli)` - links in the SDKs section - links in the Advanced usage section - links in the Telemetry section - links in the Get in touch section Additionally, this PR adds a link to the Meilisearch logo (there is currently none.) ## On the UTM pattern All links in this PR use the new convention `@gmourier` and I agreed on: - utm_campaign=oss - utm_source=github - utm_medium=meilisearch - utm_content= where the link is in the page It's worth considering updating the tracking link for the Cloud, which is the only one that doesn’t follow the new convention. It is currently using `utm_campaign=oss&utm_source=engine&utm_medium=meilisearch`. Merging analytics from different UTMs is doable on Amplitude, but can't be done in Fathom. Plus, having two different conventions creates knowledge overhead, and is bound to result in corrupt analytics at some point. I suggest we change the Cloud UTM trackers too — the sooner we eat the frog, the better imo. ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Strift <strift@Strifts-MacBook-Pro.local> Co-authored-by: Strift <laurent@meilisearch.com>	2023-07-18 13:42:50 +00:00
meili-bors[bot]	d5ab750627	Merge #3935 3935: Update mini-dashboard to version 0.2.8 r=Kerollmops a=bidoubiwa # Pull Request ## What does this PR do? - Updates the mini-dashboard to version [0.2.8](https://github.com/meilisearch/mini-dashboard/releases/tag/v0.2.8) ## PR checklist - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Charlotte Vermandel <charlottevermandel@gmail.com>	2023-07-18 12:59:29 +00:00
Charlotte Vermandel	2afd10f96d	Update mini-dashboard to version 0.2.8	2023-07-18 14:49:36 +02:00
Strift	e691c92ed5	Replace UTM link on Cloud	2023-07-18 14:48:00 +02:00
meili-bors[bot]	2d2619bd90	Merge #3933 3933: Stop computing the update files size r=ManyTheFish a=Kerollmops This PR, related #3934, removes the part which computes the total size of the `data.ms/update_files` folder, which can take a lot of time when many updates must be processed. It is not breaking API-side but is breaking on the result we will show to the user. The `databaseSize` field returned by the `/stats` endpoint will be reduced. Co-authored-by: Kerollmops <clement@meilisearch.com>	2023-07-18 12:02:08 +00:00
Kerollmops	516d2df862	Stop computing the update files size	2023-07-18 11:51:30 +02:00
meili-bors[bot]	c76b488ab1	Merge #3929 3929: Fix a panic when sorting geo fields represented by strings r=Kerollmops a=Kerollmops This issue fixes #3927 by retrieving and parsing the original string values into f64s. I also added a test to ensure we don't break it in a future version. Co-authored-by: Kerollmops <clement@meilisearch.com>	2023-07-18 09:13:22 +00:00
Kerollmops	d383afc82b	Fix the geo sort when lat and lng are strings	2023-07-17 18:28:04 +02:00
Kerollmops	f9d94c5845	Test geo sort with string lat/lng	2023-07-17 18:28:03 +02:00
Strift	928ab2f9b1	Add UTM params to contact section links	2023-07-14 18:24:03 +02:00
Strift	7c18a9375f	Add UTM params to telemetry section links	2023-07-14 18:19:46 +02:00
Strift	05a311f9be	Add UTM params to Advanced usage links	2023-07-14 18:17:51 +02:00
Strift	9b1b9b409e	Add UTM params to SDKs logos link	2023-07-14 18:17:28 +02:00
Strift	7f555f23e8	Add UTM params to SDKs section links	2023-07-14 18:15:17 +02:00
Strift	a0bfc9f63a	Add UTM params to docs & getting started links	2023-07-14 18:02:21 +02:00
Strift	3155264381	Add UTM params to features links	2023-07-14 17:51:25 +02:00
Strift	42400c381e	Add UTM on demo link	2023-07-14 17:43:05 +02:00
Strift	08c7dab528	Add UTM on demo gif	2023-07-14 17:40:37 +02:00
Strift	8590687515	Add UTM params to nav links	2023-07-14 17:34:45 +02:00
Strift	8f5d127b1e	Add links on Meilisearch logo	2023-07-14 17:26:06 +02:00
meili-bors[bot]	7745cc9d3c	Merge #3921 3921: Deactivate camel case segmentation r=dureuill a=ManyTheFish # Pull Request This PR deactivates the camel case segmentation to retrieve the possibility to accept typos over camel-cased words ## Related issue Fixes #3869 Fixes #3818 ## What does this PR do? - deactivates camelcase segmentation related to #3919 Co-authored-by: ManyTheFish <many@meilisearch.com>	2023-07-13 11:00:14 +00:00
meili-bors[bot]	657f24ec5f	Merge #3907 3907: Add telemetry for define field to search on at query time r=dureuill a=ManyTheFish Add "attributes_to_search_on" telemetry usage counter: ```json "attributes_to_search_on": { "total_number_of_use": 12, }, ``` This measures the number of search queries that the user uses `attributesToSearchOn` field. related to https://github.com/meilisearch/specifications/pull/251 ## reviewers: - `@macraig` for validating the telemetry's name - `@dureuill` for validating the code Co-authored-by: ManyTheFish <many@meilisearch.com>	2023-07-13 10:14:00 +00:00
ManyTheFish	c106906f8f	deactivate camelCase segmentation	2023-07-13 12:06:27 +02:00
ManyTheFish	9c0691156f	Add tests	2023-07-13 11:53:13 +02:00
ManyTheFish	359b90288d	Use saturating add	2023-07-13 11:38:28 +02:00
ManyTheFish	13e3f8faae	Fix typo	2023-07-13 11:34:50 +02:00
meili-bors[bot]	fd7c66fd62	Merge #3915 3915: `attributesToSearchOn` supports wildcards r=ManyTheFish a=dureuill # Pull Request ## Related issue Fixes #3912 and #3911 ## What does this PR do? - Adding `` in the list of `attributesToSearchOn` allows searching on all the `searchableAttributes`. - If `searchableAttributes contains ""`, then any attribute is accepted in the `attributesToSearchOn` list. Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-07-13 09:33:10 +00:00
Louis Dureuil	183f23f40d	More relevant test Co-authored-by: Many the fish <many@meilisearch.com>	2023-07-12 16:06:15 +02:00
meili-bors[bot]	2b4160ebb9	Merge #3918 3918: Update and fix the Test Suite CI r=dureuill a=Kerollmops This Pull Request renames the _Run test with Rust_ into _Setup test with Rust_ for more clarity and `cargo update -p proc-macro2` to make the project compile with the latest Rust Nightly. Co-authored-by: Kerollmops <clement@meilisearch.com>	2023-07-12 13:18:25 +00:00
Kerollmops	8ba1c8f88f	Update proc-macro2 to compile with the latest nightly	2023-07-12 11:47:27 +02:00
Louis Dureuil	16c8437b28	Update tests	2023-07-12 11:21:19 +02:00
Kerollmops	8e7edf8ea7	Rename the jobs in the CI for clarity	2023-07-12 11:16:01 +02:00
Louis Dureuil	4310928803	Fixes #3912	2023-07-12 10:08:56 +02:00
Louis Dureuil	74315b4ea8	Fixes #3911	2023-07-12 10:08:29 +02:00
meili-bors[bot]	177e6e27f9	Merge #3901 3901: Fix experimental analytics r=curquiza a=dureuill # Pull Request ## Related issue Fixes https://github.com/meilisearch/specifications/pull/250#discussion_r1253191583 ## What does this PR do? - `snake_case` instead of `camelCase` for feature fields Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-07-10 16:22:59 +00:00
meili-bors[bot]	50afe724ae	Merge #3909 3909: Effectively send the `vector.max_vector_size` telemetry r=curquiza a=Kerollmops This PR effectively aggregates and sends the `vector.max_vector_size` analytics value. Co-authored-by: Kerollmops <clement@meilisearch.com>	2023-07-10 15:44:30 +00:00
Kerollmops	012c960fad	Send the vector.max_vector_size telemetry	2023-07-10 16:50:37 +02:00
meili-bors[bot]	76f6d3357e	Merge #3908 3908: Allow a comma-separated value to the `vector` argument in GET search r=Kerollmops a=dureuill # Pull Request For request: ``` curl \ -X GET 'http://localhost:7700/indexes/movies/search?vector=0.123,1.124,244' ``` Before PR: ``` {"message":"Invalid value type for parameter `vector`: expected a string, but found a string: `0,1,2`","code":"invalid_search_vector","type":"invalid_request","link":"https://docs.meilisearch.com/errors#invalid_search_vector"}% ``` After PR: ``` {"hits":[],"query":"","vector":[0.123,1.124,244.0],"processingTimeMs":0,"limit":20,"offset":0,"estimatedTotalHits":1000}% ``` cc `@gmourier` `@bidoubiwa` Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-07-10 14:25:44 +00:00
Louis Dureuil	d59e969c16	Allow a comma-separated value to the `vector` argument in GET search	2023-07-10 16:16:34 +02:00
meili-bors[bot]	eb7a1aa7af	Merge #3904 3904: Sort by lexicographic order after normalization r=dureuill a=dureuill # Pull Request ## Related issue Fixes https://github.com/meilisearch/meilisearch/issues/3893 ## What does this PR do? - Re-sort stop words after normalization so they're not sent out-of-order to the FST Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-07-10 12:12:05 +00:00
meili-bors[bot]	9daccdf7f0	Merge #3895 3895: Update README.md r=curquiza a=ferdi05 Adding the free-trial option # Pull Request ## Related issue Fixes #<issue_number> ## What does this PR do? - ... ## PR checklist Please check if your PR fulfills the following requirements: - [ ] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [ ] Have you read the contributing guidelines? - [ ] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Ferdinand Boas <ferdinand.boas@gmail.com>	2023-07-10 11:26:47 +00:00
ManyTheFish	c30a14cb97	Add telemetry	2023-07-10 13:12:12 +02:00
meili-bors[bot]	a3ca8412ce	Merge #3906 3906: Add "scoring.*" analytics to multi search route r=Kerollmops a=dureuill # Pull Request ## Related issue Fixes https://github.com/meilisearch/specifications/pull/252#discussion_r1254375746 by implementing (3): multi search now returns the "score.show_ranking_rule" and "score.show_ranking_rule_details" analytics. Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-07-10 09:51:30 +00:00
Louis Dureuil	106f98aa72	Add "scoring.*" analytics to multi search route	2023-07-10 11:45:43 +02:00
Louis Dureuil	40fa59d64c	Sort by lexicographic order after normalization	2023-07-10 09:26:59 +02:00
Louis Dureuil	bb40ce6e35	Experimental features analytics match the spec	2023-07-10 08:57:53 +02:00
meili-bors[bot]	0c8dbf6fa6	Merge #3897 3897: Add automated tests for `/experimental-features` route r=Kerollmops a=dureuill # Pull Request ## What does this PR do? - Make `RuntimeTogglableFeatures` `Eq` - Add various tests for the `/experimental-features` route - Integration tests for the route itself - Integration tests for the effect of enabling `scoreDetails` and `vectorStore` through this route. - Dump integration tests Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-07-06 13:37:56 +00:00
Louis Dureuil	dd6519b64f	Dump tests	2023-07-06 14:22:29 +02:00
Louis Dureuil	da02a9cf32	Make RuntimeTogglableFeatures Eq	2023-07-06 14:20:58 +02:00
meili-bors[bot]	ff192bb480	Merge #3889 3889: Display the total number of tasks matching a filter/query r=dureuill a=Kerollmops This PR returns a new field on the `/tasks` routes. The `total` field exposes the total number of tasks that matches the given filter/query. It is useful to display information on a user interface and can help understand when progress is made in processing tasks, i.e., the total number of tasks on `/tasks?statuses=succeeded` will increase over time. Fixes #3888. - [ ] Update the specs fo the `/tasks` route. ## How have I implemented it? I found it much easier to run two times the task filtering system. Once with the original `from` and `limit` parameters and a second time without. The second call will return the total number of tasks that match the query, not only the number of tasks on the current page. So far, in terms of performance, there doesn't seem to be any issue. I tried different filters with something like 250k tasks. Note that there is a limit of 1M tasks in the queue. Co-authored-by: Clément Renault <clement@meilisearch.com>	2023-07-06 10:23:09 +00:00
Ferdinand Boas	437ee55c57	Update README.md Co-authored-by: Guillaume Mourier <guillaume@meilisearch.com>	2023-07-06 12:15:52 +02:00
Clément Renault	22762808ab	Fix the tests	2023-07-06 12:13:29 +02:00
Ferdinand Boas	b1717865ea	Update README.md Adding the free-trial option	2023-07-06 11:52:35 +02:00
Clément Renault	86b834c9e4	Display the total number of tasks in the tasks route	2023-07-06 10:05:18 +02:00
Louis Dureuil	2d3cec11a7	Search integration test to check score details and vector store	2023-07-06 09:02:02 +02:00
Louis Dureuil	76e1ee9988	integration test on "/experimental-features" route	2023-07-06 09:01:28 +02:00
Louis Dureuil	222615d3df	Allow to get/set features in integration test server	2023-07-06 09:01:05 +02:00
Louis Dureuil	11d024c613	Authentication tests	2023-07-06 09:00:51 +02:00
meili-bors[bot]	886c8bb647	Merge #3891 3891: Fix the way we compute the 99th percentile r=dureuill a=Kerollmops This PR fixes how we compute the 99th percentile by avoiding using float and doing the multiplication and divisions in the correct order avoiding going out of the buffer of timings. You can see the issue on [this rust playground](https://play.rust-lang.org/?version=stable&mode=debug&edition=2021). When there are a very small number of successful requests, the number is so tiny that the 99th percentile calculus sometimes gives an index out of the buffer. In this example, the `1`/`1.0` represent the number of timings you collected (one). As you can see, the float computation gives us the index `1.0`, with is out of a vector of only one value. This makes the engine generate a `null` value. ```rust 1 * 99 / 100 = 0 // with integers 0.99_f64 * (1.0 - 1.0) + 1.0 = 1.0 // with floats ``` Co-authored-by: Clément Renault <clement@meilisearch.com>	2023-07-06 06:04:08 +00:00
meili-bors[bot]	b422e5fdc3	Merge #3890 3890: Fix the analytics of the sort facet values by count feature r=dureuill a=Kerollmops This PR ensures we return the right analytics from the settings route. Co-authored-by: Clément Renault <clement@meilisearch.com>	2023-07-06 05:24:40 +00:00
Clément Renault	d727ebee05	Fix the way we compute the 99th percentile	2023-07-05 17:53:09 +02:00
Clément Renault	da39a7b29e	Return the right analytics	2023-07-05 17:27:51 +02:00
meili-bors[bot]	377fe33aac	Merge #3885 3885: Exactness missing field r=dureuill a=dureuill # Pull Request Adds fields to score details that were [specified](`c25d758264/text/0195-ranking-score.md (322-ranking-rule-specific-fields)`), but missing in the implementation: - `exactness.matchingWords` - `exactness.maxMatchingWords` Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-07-04 15:14:53 +00:00
Louis Dureuil	55cd7738b9	Update snapshots	2023-07-04 16:31:01 +02:00
Louis Dureuil	48409c9183	Add missing exactness.matchingWords, exactness.maxMatchingWords	2023-07-04 16:31:01 +02:00
meili-bors[bot]	176f716292	Merge #3871 3871: Bump Swatinem/rust-cache from 2.4.0 to 2.5.0 r=curquiza a=dependabot[bot] Bumps [Swatinem/rust-cache](https://github.com/swatinem/rust-cache) from 2.4.0 to 2.5.0. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/swatinem/rust-cache/releases">Swatinem/rust-cache's releases</a>.</em></p> <blockquote> <h2>v2.5.0</h2> <h2>What's Changed</h2> <ul> <li>feat: Rm workspace crates version before caching by <a href="https://github.com/NobodyXu"><code>`@NobodyXu</code></a>` in <a href="https://redirect.github.com/Swatinem/rust-cache/pull/147">Swatinem/rust-cache#147</a></li> <li>feat: Add hash of <code>.cargo/config.toml</code> to key by <a href="https://github.com/NobodyXu"><code>`@NobodyXu</code></a>` in <a href="https://redirect.github.com/Swatinem/rust-cache/pull/149">Swatinem/rust-cache#149</a></li> </ul> <h2>New Contributors</h2> <ul> <li><a href="https://github.com/NobodyXu"><code>`@NobodyXu</code></a>` made their first contribution in <a href="https://redirect.github.com/Swatinem/rust-cache/pull/147">Swatinem/rust-cache#147</a></li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/Swatinem/rust-cache/compare/v2.4.0...v2.5.0">https://github.com/Swatinem/rust-cache/compare/v2.4.0...v2.5.0</a></p> </blockquote> </details> <details> <summary>Changelog</summary> <p><em>Sourced from <a href="https://github.com/Swatinem/rust-cache/blob/master/CHANGELOG.md">Swatinem/rust-cache's changelog</a>.</em></p> <blockquote> <h1>Changelog</h1> </blockquote> </details> <details> <summary>Commits</summary> <ul> <li><a href="`2656b87321`"><code>2656b87</code></a> 2.5.0</li> <li><a href="`715970feed`"><code>715970f</code></a> feat: Add hash of <code>.cargo/config.toml</code> to key (<a href="https://redirect.github.com/swatinem/rust-cache/issues/149">#149</a>)</li> <li><a href="`3d4000164d`"><code>3d40001</code></a> feat: Rm workspace crates version before caching (<a href="https://redirect.github.com/swatinem/rust-cache/issues/147">#147</a>)</li> <li>See full diff in <a href="https://github.com/swatinem/rust-cache/compare/v2.4.0...v2.5.0">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=Swatinem/rust-cache&package-manager=github_actions&previous-version=2.4.0&new-version=2.5.0)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) You can trigger a rebase of this PR by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) </details> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2023-07-04 12:57:38 +00:00
meili-bors[bot]	82650eaae1	Merge #3877 3877: update the total_received properties of multiple events r=dureuill a=dureuill # Pull Request ## Related issue Fixes #3814 ## What does this PR do? -fix name of `total_received` for several events Co-authored-by: Tamo <tamo@meilisearch.com>	2023-07-03 19:49:53 +00:00
meili-bors[bot]	b8ca09c13f	Merge #3878 3878: Remove unsafe `atty` dependency r=dureuill a=Kerollmops This PR replaces the `atty` dependency with the `is-terminal` one. We do that to fix GHSA-g98v-hv3f-hcfr. Co-authored-by: Kerollmops <clement@meilisearch.com>	2023-07-03 19:07:03 +00:00
Kerollmops	a442af6a7c	Update the features of the either dependency to compile milli successfully	2023-07-03 18:51:43 +02:00
Kerollmops	e7f8daaf86	Update criterion to 0.5.1 to remove the atty dependency	2023-07-03 18:51:42 +02:00
Kerollmops	d1ff631df8	Replace the atty dependency with the is-terminal one	2023-07-03 18:51:42 +02:00
Tamo	202183adf8	update the total_received properties of multiple events	2023-07-03 15:57:09 +02:00
meili-bors[bot]	aae099e330	Merge #3851 3851: Expose lastUpdate and isIndexing in /stats endpoint r=dureuill a=gentcys # Pull Request ## Related issue Fixes #3843 ## What does this PR do? - expose lastUpdate in `/stats` endpoint - expose isIndex in `stats` endpoint - add a method `is_task_processing` in index-scheduler/src/lib.rs. ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Cong Chen <cong.chen@ocrlabs.com> Co-authored-by: ManyTheFish <many@meilisearch.com> Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-07-03 13:41:04 +00:00
Louis Dureuil	5387cf1718	Don't unwrap in case of error/missing last_update field	2023-07-03 15:32:11 +02:00
meili-bors[bot]	a0df4becf4	Merge #3867 3867: Add a new link to the cloud pricing page r=curquiza a=Kerollmops This PR promotes the Cloud by adding a link to the Pricing page to the startup message! <img width="1002" alt="Capture d’écran 2023-06-29 à 17 40 22" src="https://github.com/meilisearch/meilisearch/assets/3610253/b0528c24-fcc2-43ff-a6a1-3ed91716663b"> Co-authored-by: Clément Renault <clement@meilisearch.com>	2023-07-03 11:25:26 +00:00
meili-bors[bot]	e0a2f88fb0	Merge #3874 3874: Update version for the next release (v1.3.0) in Cargo.toml r=curquiza a=meili-bot ⚠️ This PR is automatically generated. Check the new version is the expected one and Cargo.lock has been updated before merging. Co-authored-by: gillian-meilisearch <gillian-meilisearch@users.noreply.github.com>	2023-07-03 10:37:03 +00:00
meili-bors[bot]	e871906370	Merge #3876 3876: Fix invalid attributeToSearchOn error code r=Kerollmops a=ManyTheFish Fix the invalid attributeToSearchOn error code to be consistent with the other search parameters' error codes: error code `invalid_attributes_to_search_on` becomes `invalid_search_attributes_to_search_on`: ```diff - invalid_attributes_to_search_on + invalid_search_attributes_to_search_on ``` related to #3772 Co-authored-by: ManyTheFish <many@meilisearch.com>	2023-07-03 10:06:30 +00:00
ManyTheFish	7a80c0dfb3	Fix invalid attributeToSearchOn error code to be consistent with the others search parameters error codes	2023-07-03 11:52:43 +02:00
ManyTheFish	71500a4e15	Update tests	2023-07-03 11:20:43 +02:00
meili-bors[bot]	a9f691f279	Merge #3873 3873: Format let-else ❤️ 🎉 r=Kerollmops a=dureuill # Pull Request Allows passing CI after landing of `6162f6f123` ## What does this PR do? - `cargo +nightly fmt` Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-07-03 09:01:20 +00:00
gillian-meilisearch	1d40452057	Update version for the next release (v1.3.0) in Cargo.toml	2023-07-03 08:32:21 +00:00
Louis Dureuil	324d448236	Format let-else ❤️ 🎉	2023-07-03 10:20:28 +02:00
dependabot[bot]	40ad19ba9e	Bump Swatinem/rust-cache from 2.4.0 to 2.5.0 Bumps [Swatinem/rust-cache](https://github.com/swatinem/rust-cache) from 2.4.0 to 2.5.0. - [Release notes](https://github.com/swatinem/rust-cache/releases) - [Changelog](https://github.com/Swatinem/rust-cache/blob/master/CHANGELOG.md) - [Commits](https://github.com/swatinem/rust-cache/compare/v2.4.0...v2.5.0) --- updated-dependencies: - dependency-name: Swatinem/rust-cache dependency-type: direct:production update-type: version-update:semver-minor ... Signed-off-by: dependabot[bot] <support@github.com>	2023-07-01 17:46:11 +00:00
Cong Chen	9859e65d2f	fix tests	2023-07-01 09:32:50 +08:00
Cong Chen	3bdf01bc1c	Fix failed test	2023-06-30 17:39:23 +08:00
Cong Chen	a5a31667b0	fix converse result of is_task_processing()	2023-06-30 11:28:18 +08:00
Clément Renault	cab4c4d7c9	Add a UTMs to the Cloud link	2023-06-29 17:59:59 +02:00
Clément Renault	4ec08e9430	Add a new link to the cloud pricing page	2023-06-29 17:38:10 +02:00
meili-bors[bot]	661d1f90dc	Merge #3866 3866: Update charabia v0.8.0 r=dureuill a=ManyTheFish # Pull Request Update Charabia: - enhance Japanese segmentation - enhance Latin Tokenization - words containing `_` are now properly segmented into several words - brackets `{([])}` are no more considered as context separators so word separated by brackets are now considered near together for the proximity ranking rule - fixes #3815 - fixes #3778 - fixes [product#151](https://github.com/meilisearch/product/discussions/151) > Important note: now the float numbers are segmented around the `.` so `3.22` is segmented as [`3`, `.`, `22`] but the middle dot isn't considered as a hard separator, which means that if we search `3.22` we find documents containing `3.22` Co-authored-by: ManyTheFish <many@meilisearch.com>	2023-06-29 15:24:36 +00:00
ManyTheFish	6ec7541026	Update inta snapshots	2023-06-29 17:18:39 +02:00
ManyTheFish	e8dee3ca65	Update lock file	2023-06-29 17:02:24 +02:00
ManyTheFish	a82c49ab08	Update test	2023-06-29 15:56:36 +02:00
ManyTheFish	84845de9ef	Update Charabia	2023-06-29 15:56:32 +02:00
meili-bors[bot]	c9b3f80947	Merge #3780 3780: Be able to sort facet values by alpha or count r=dureuill a=Kerollmops This PR introduces a new `sortFacetValuesBy` settings parameter to expose the facet distribution in either count or lexicographic/alpha order. ## Mini Spec of the `sortFacetValuesBy` Settings Parameter This parameter can be set in the settings to change how the engine returns the facet values. There are two possible values to this parameter. Please note that the current behavior changed a bit, and keys are returned in lexicographic order instead of undefined order. The previous order wasn't defined as we were using a `HashMap`, which returns entries in hash order (undefined), and we are now using an `IndexMap`, which returns them in insertion order (the order we actually want). Also, note that there are performance issues when the dataset is enormous. Here are the timings of the engine running on my Macbook Pro M1 (16Go of RAM). [The dataset is 40 million songs file](https://www.notion.so/meilisearch/Songs-from-MusicBrainz-686e31b2bd3845898c7746f502a6e117), and the database size is about 50GiB. Even if you think 800ms is not that high, don't forget that the API is public, and anybody can ask for multiple facets in a single query. \| Search Kind \| Get Facets \| Max Values per Facet \| Time for Alpha \| Time for Count \| Count but with #3788 \| \|------------:\|------------\|----------------------\|:--------------:\|----------------\|----------------------\| \| Placeholder \| genres \| default (100) \| 7ms \| 187ms \| 122ms \| \| Placeholder \| genres \| 20 \| 6ms \| 124ms \| 75ms \| \| Placeholder \| album \| default (100) \| 9ms \| 808ms \| 677ms \| \| Placeholder \| album \| 20 \| 8ms \| 579ms \| 446ms \| \| Placeholder \| artist \| default (100) \| 9ms \| 462ms \| 344ms \| \| Placeholder \| artist \| 20 \| 9ms \| 341ms \| 246ms \| ### Order Values in Alphanumeric Order This is the default one. Values will be returned by lexicographic order, ascending from A to Z. ```bash # First, update the settings curl 'localhost:7700/indexes/movies/settings/facetting' \ -H "Content-Type: application/json" \ -d '{ "sortFacetValuesBy": { "": "alpha" } }' # Then, ask for the facet distribution curl 'localhost:7700/indexes/movies/search?facets=genres' ``` ```json5 { "hits": [ / list of results / ], "query": "", "processingTimeMs": 0, "limit": 20, "offset": 0, "estimatedTotalHits": 1000, "facetDistribution": { "genres": { "Action": 3215, "Adventure": 1972, "Animation": 1577, "Comedy": 5883, "Crime": 1808, // ... } }, "facetStats": {} } ``` ### Order Values in Count Order Facet values are sorted by decreasing count. The count is the number of records containing this facet value in the query results. ```bash # First, update the settings curl 'localhost:7700/indexes/movies/settings/facetting' \ -H "Content-Type: application/json" \ -d '{ "sortFacetValuesBy": { "": "count" } }' # Then, ask for the facet distribution curl 'localhost:7700/indexes/movies/search?facets=genres' ``` ```json5 { "hits": [ /* list of results */ ], "query": "", "processingTimeMs": 0, "limit": 20, "offset": 0, "estimatedTotalHits": 1000, "facetDistribution": { "genres": { "Drama": 7337, "Comedy": 5883, "Action": 3215, "Thriller": 3189, "Romance": 2507, // ... } }, "facetStats": {} } ``` ## Todo List - [x] Add tests - [x] Send analytics when a user change the `sortFacetValuesBy` - [x] Create a prototype and announce it in https://github.com/meilisearch/product/discussions/519. Co-authored-by: Kerollmops <clement@meilisearch.com> Co-authored-by: Clément Renault <clement@meilisearch.com>	2023-06-29 12:43:25 +00:00
Clément Renault	09c5edf242	Cargo fmt	2023-06-29 14:37:18 +02:00
Clément Renault	4e85f91aee	Add a non default value to the faceting settings of the dump tests	2023-06-29 14:33:33 +02:00
Clément Renault	7c157fc442	Document that the LevelEntry fields order is important	2023-06-29 14:33:32 +02:00
Clément Renault	0b97596c93	Replace unwraps with ?	2023-06-29 14:33:32 +02:00
Clément Renault	a0e0fce677	Simplify a Rust lifetime trick	2023-06-29 14:33:32 +02:00
Clément Renault	3c295c1ffc	Fix typos	2023-06-29 14:33:32 +02:00
Clément Renault	b951830461	Add more tests	2023-06-29 14:33:32 +02:00
Clément Renault	9a13b72f25	Fix the tests	2023-06-29 14:33:32 +02:00
Clément Renault	1d8dfafd25	Add analytics when all facets are sorted by count and the number of modified ones	2023-06-29 14:33:31 +02:00
Kerollmops	eed9176e0c	Also reset the sortFacetValuesBy when reseting the faceting settings	2023-06-29 14:33:31 +02:00
Kerollmops	b132e859f7	Make clippy happy	2023-06-29 14:33:31 +02:00
Kerollmops	9917bf046a	Move the sortFacetValuesBy in the faceting settings	2023-06-29 14:33:31 +02:00
Kerollmops	d9fea0143f	Make Clippy happy	2023-06-29 14:33:31 +02:00
Kerollmops	a385642ec3	Replace the BTreeMap by an IndexMap to return values in order	2023-06-29 14:33:31 +02:00
Kerollmops	34b2e98fe9	Expose a sortFacetValuesBy parameter to the user	2023-06-29 14:33:00 +02:00
Kerollmops	80bbd4b6f3	Clean and make the facet order configurable internally	2023-06-29 14:31:17 +02:00
Kerollmops	f42bef2f66	Make the search to always return the facets ordered by count	2023-06-29 14:31:17 +02:00
Kerollmops	bd3c026406	First to-test version of the algorithm	2023-06-29 14:31:17 +02:00
Kerollmops	84f8938f33	Rename facet distribution to be explicit on the order to find them	2023-06-29 14:31:15 +02:00
meili-bors[bot]	34a07110de	Merge #3864 3864: Remove `/experimental-features` verbs that weren't in the PRD r=dureuill a=dureuill Removes: - POST `/experimental-features` - DELETE `/experimental-features` keeping only: - PATCH `/experimental-features` - GET `/experimental-features` The two routes that are described in the PRD. Following `@guimachiavelli's` [question](https://github.com/meilisearch/documentation/issues/2482#issuecomment-1611845372) about the POST route. Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-06-29 09:43:14 +00:00
meili-bors[bot]	73bb080a26	Merge #3699 3699: Search for Facet Values r=Kerollmops a=Kerollmops This PR introduces the first version of [the _Search for Facet Values_ feature](https://github.com/meilisearch/product/discussions/515) that allows a user to search for facets, by optionally using a prefix string and optionally specifying the `q` and `filter` original search parameters to restrict the candidates to search in. The steps to merge it into Meilisearch will first start by providing prototype Docker images. This way users will be able to test the prototypes before using them. The current route to use the _Search for Facet Values_ feature is the `POST /indexes/{index}/facet-search` where the body is a JSON object that looks like the following: ```json5 { "q": "spiderman", // optional "filter": "rating > 10", // optional "facetName": "genres", "facetQuery": "a" // optional } ``` ## What is missing? - [x] Send some analytics. - [x] Support the `matchingStrategy` parameter. - [x] Make sure that the errors are the right ones. - [x] Use the [Index typo tolerance settings](https://www.meilisearch.com/docs/learn/configuration/typo_tolerance#minwordsizefortypos) when matching facet values. - [x] minWordSizeForTypos.oneTypo - [x] minWordSizeForTypos.twoTypo - [x] Add tests - [x] Log the time it took to compute the results. - [x] Fix the compilation warnings. - [x] [Create an issue to fix potential performance issues when indexing](https://github.com/meilisearch/meilisearch/issues/3862). Co-authored-by: Clément Renault <clement@meilisearch.com> Co-authored-by: Kerollmops <clement@meilisearch.com>	2023-06-29 09:08:55 +00:00
Clément Renault	44b5b9e1a7	Improve the documentation of the FacetSearchQuery struct	2023-06-29 10:28:23 +02:00
Louis Dureuil	68356869c0	Remove `/experimental-features` verbs that weren't in the PRD	2023-06-29 10:02:55 +02:00
Cong Chen	e3fc7112bc	use `RoaringBitmap::is_empty` instead	2023-06-29 11:46:47 +08:00
Louis Dureuil	605c1dd54a	Fix analytics	2023-06-28 16:41:56 +02:00
Clément Renault	3e3f73ba1e	Fix the analytics	2023-06-28 15:45:09 +02:00
Clément Renault	efbe7ce78b	Clean the facet string FSTs when we clear the documents	2023-06-28 15:36:32 +02:00
Louis Dureuil	82e1f59f1e	Add attributes_to_search_on	2023-06-28 15:28:24 +02:00
Clément Renault	362e9ff845	Add more tests	2023-06-28 15:28:24 +02:00
Clément Renault	32f2556d22	Move the additional_search_parameters_provided analytic inside facets	2023-06-28 15:06:09 +02:00
Kerollmops	63fd10aaa5	Fix the invalid facet name field error code	2023-06-28 15:06:09 +02:00
Kerollmops	29b40295b8	Ignore unknown facet search query parameters	2023-06-28 15:06:09 +02:00
Kerollmops	26f0fa678d	Change the error message when a facet is not searchable	2023-06-28 15:06:09 +02:00
Kerollmops	60ddd53439	Return one of the original facet values when doing a facet search	2023-06-28 15:06:09 +02:00
Kerollmops	2bcd8d2983	Make sure the facet queries are normalized	2023-06-28 15:06:09 +02:00
Kerollmops	09079a4e88	Remove useless InvalidSearchFacet error	2023-06-28 15:06:09 +02:00
Kerollmops	904f6574bf	Make rustfmt happy	2023-06-28 15:06:08 +02:00
Kerollmops	6fb8af423c	Rename the hits and query output into facetHits and facetQuery respectively	2023-06-28 15:06:08 +02:00
Kerollmops	cb0bb399fa	Fix the error code returned when the facetName field is missing	2023-06-28 15:06:08 +02:00
Kerollmops	41760a9306	Introduce a new invalid_facet_search_facet_name error code	2023-06-28 15:06:07 +02:00
Kerollmops	e9a3029c30	Use the right field id to write the string facet values FST	2023-06-28 15:01:51 +02:00
Kerollmops	ed0ff47551	Return an empty list of results if attribute is set as filterable	2023-06-28 15:01:51 +02:00
Clément Renault	e1b8fb48ee	Use the minWordSizeForTypos index settings	2023-06-28 15:01:51 +02:00
Clément Renault	87e22e436a	Fix compilation issues	2023-06-28 15:01:51 +02:00
Clément Renault	0252cfe8b6	Simplify the placeholder search of the facet-search route	2023-06-28 15:01:50 +02:00
Clément Renault	f35ad96afa	Use the disableOnAttributes parameter on the facet-search route	2023-06-28 15:01:50 +02:00
Clément Renault	2ceb781c73	Use the disableOnWords parameter on the facet-search route	2023-06-28 15:01:50 +02:00
Clément Renault	7bd67543dd	Support the typoTolerant.enabled parameter	2023-06-28 15:01:50 +02:00
Clément Renault	8e86eb91bb	Log an error when a facet value is missing from the database	2023-06-28 15:01:50 +02:00
Clément Renault	55c17aa38b	Rename the SearchForFacetValues struct	2023-06-28 15:01:50 +02:00
Clément Renault	aadbe88048	Return an internal error when a field id is missing	2023-06-28 15:01:50 +02:00
Clément Renault	f36de2115f	Make clippy happy	2023-06-28 15:01:50 +02:00
Clément Renault	702041b7e1	Improve the returned errors from the facet-search route	2023-06-28 15:01:48 +02:00
Clément Renault	a05074e675	Fix the max number of facets to be returned to 100	2023-06-28 14:58:42 +02:00
Clément Renault	93f30e65a9	Return the correct response JSON object from the facet-search route	2023-06-28 14:58:42 +02:00
Clément Renault	893592c5e9	Send analytics about the facet-search route	2023-06-28 14:58:42 +02:00
Clément Renault	e81809aae7	Make the search for facet work	2023-06-28 14:58:41 +02:00
Kerollmops	ce7e7f12c8	Introduce the facet search route	2023-06-28 14:58:41 +02:00
Kerollmops	addb21f110	Restrict the number of facet search results to 1000	2023-06-28 14:58:41 +02:00
Kerollmops	c34de05106	Introduce the SearchForFacetValue struct	2023-06-28 14:58:41 +02:00
Clément Renault	15a4c05379	Store the facet string values in multiple FSTs	2023-06-28 14:58:41 +02:00
meili-bors[bot]	9deeec88e0	Merge #3861 3861: Add "meilisearch" prefix to last metrics that were missing it r=Kerollmops a=dureuill # Pull Request ## Related issue Related to #3790 ## What does this PR do? - change implementation to follow the spec on metrics name - regenerate grafana dashboard from the code ## PR checklist Please check if your PR fulfills the following requirements: - [ ] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [ ] Have you read the contributing guidelines? - [ ] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-06-28 09:28:31 +00:00
Louis Dureuil	167ac55a2d	Update dashboard generated from grafana	2023-06-28 11:22:16 +02:00
Louis Dureuil	ea68ccd034	prefix http_* metrics by meilisearch	2023-06-28 11:21:50 +02:00
meili-bors[bot]	d4f10800f2	Merge #3834 3834: Define searchable fields at runtime r=Kerollmops a=ManyTheFish ## Summary This feature allows the end-user to search in one or multiple attributes using the search parameter `attributesToSearchOn`: ```json { "q": "Captain Marvel", "attributesToSearchOn": ["title"] } ``` This feature act like a filter, forcing Meilisearch to only return the documents containing the requested words in the attributes-to-search-on. Note that, with the matching strategy `last`, Meilisearch will only ensure that the first word is in the attributes-to-search-on, but, the retrieved documents will be ordered taking into account the word contained in the attributes-to-search-on. ## Trying the prototype A dedicated docker image has been released for this feature: #### last prototype version: ```bash docker pull getmeili/meilisearch:prototype-define-searchable-fields-at-search-time-1 ``` #### others prototype versions: ```bash docker pull getmeili/meilisearch:prototype-define-searchable-fields-at-search-time-0 ``` ## Technical Detail The attributes-to-search-on list is given to the search context, then, the search context uses the `fid_word_docids`database using only the allowed field ids instead of the global `word_docids` database. This is the same for the prefix databases. The database cache is updated with the merged values, meaning that the union of the field-id-database values is only made if the requested key is missing from the cache. ### Relevancy limits Almost all ranking rules behave as expected when ordering the documents. Only `proximity` could miss-order documents if all the searched words are in the restricted attribute but a better proximity is found in an ignored attribute in a document that should be ranked lower. I put below a failing test showing it: ```rust #[actix_rt::test] async fn proximity_ranking_rule_order() { let server = Server::new().await; let index = index_with_documents( &server, &json!([ { "title": "Captain super mega cool. A Marvel story", // Perfect distance between words in an ignored attribute "desc": "Captain Marvel", "id": "1", }, { "title": "Captain America from Marvel", "desc": "a Shazam ersatz", "id": "2", }]), ) .await; // Document 2 should appear before document 1. index .search(json!({"q": "Captain Marvel", "attributesToSearchOn": ["title"], "attributesToRetrieve": ["id"]}), \|response, code\| { assert_eq!(code, 200, "{}", response); assert_eq!( response["hits"], json!([ {"id": "2"}, {"id": "1"}, ]) ); }) .await; } ``` Fixing this would force us to create a `fid_word_pair_proximity_docids` and a `fid_word_prefix_pair_proximity_docids` databases which may multiply the keys of `word_pair_proximity_docids` and `word_prefix_pair_proximity_docids` by the number of attributes in the searchable_attributes list. If we think we should fix this test, I'll suggest doing it in another PR. ## Related Fixes #3772 Co-authored-by: Tamo <tamo@meilisearch.com> Co-authored-by: ManyTheFish <many@meilisearch.com>	2023-06-28 08:19:23 +00:00
meili-bors[bot]	dc293911ad	Merge #3745 3745: tests: add unit test for `PayloadTooLarge` error r=curquiza a=cymruu # Pull Request Add a unit test for the `Payload`, which verifies that a request with a payload that is too large is rejected with the appropriate message. This was requested in this PR https://github.com/meilisearch/meilisearch/pull/3739 ## Related issue https://github.com/meilisearch/meilisearch/pull/3739 ## What does this PR do? - Adds requested test ## PR checklist Please check if your PR fulfills the following requirements: - [ ] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [ ] Have you read the contributing guidelines? - [ ] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Filip Bachul <filipbachul@gmail.com>	2023-06-27 14:58:23 +00:00
meili-bors[bot]	9d68e6969e	Merge #3859 3859: Merge all analytics events pertaining to updating the experimental features r=Kerollmops a=dureuill Follow-up to #3850 Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-06-27 13:26:01 +00:00
Louis Dureuil	b4b686d253	Merge all analytics events pertaining to updating the experimental features	2023-06-27 15:16:23 +02:00
meili-bors[bot]	98ec476198	Merge #3855 3855: Change and add links to the Cloud r=Kerollmops a=dureuill - add cloud link in banner - add utm to existing links following https://github.com/meilisearch/integration-guides/issues/277#issuecomment-1592054536 Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-06-27 12:29:36 +00:00
Louis Dureuil	c47b8a8bfe	Fix typo Co-authored-by: Guillaume Mourier <guillaume@meilisearch.com>	2023-06-27 14:27:54 +02:00
Louis Dureuil	054f81a021	Make message consistent with the one in integration repos	2023-06-27 14:20:55 +02:00
meili-bors[bot]	d8ea688481	Merge #3825 3825: Accept semantic vectors and allow users to query nearest neighbors r=Kerollmops a=Kerollmops This Pull Request brings a new feature to the current API. The engine accepts a new `_vectors` field akin to the `_geo` one. This vector is stored in Meilisearch and can be retrieved via search. This work is the first step toward hybrid search, bringing the best of both worlds: keyword and semantic search ❤️‍🔥 ## ToDo - [x] Make it possible to get the `limit` nearest neighbors from a user-generated vector by using the `vector` field of search route. - [x] Delete the documents and vectors from the HNSW-related data structures. - [x] Do it the slow and ugly way (we need to be able to iterate over all the values). - [ ] Do it the efficient way (Wait for a new method or implement it myself). - [ ] ~~Move from the `hnsw` crate to the hgg one~~ The hgg crate is too slow. Meilisearch takes approximately 88s to answer a query. It is related to the time it takes to deserialize the `Hgg` data structure or search in it. I didn't take the time to measure precisely. We moved back to the hnsw crate which takes approximately 40ms to answer. - [ ] ~~Wait for a fix for https://github.com/rust-cv/hgg/issues/4.~~ - [x] Fix the current dot product function. - [x] Fill in the other `SearchResult` fields. - [x] Remove the `hnsw` dependency of the meilisearch crate. - [x] Fix the pages by taking the offset into account. - [x] Release a first prototype https://github.com/meilisearch/product/discussions/621#discussioncomment-6183647 - [x] Make the pagination and filtering faster and more correct. - [x] Return the original vector in the output search results (like `query`). - [x] Return an `_semanticSimilarity` field in the documents (it's a dot product) - [x] Return this score even if the `_vectors` field is not displayed - [x] Rename the field `_semanticScore`. - [ ] Return the `_geoDistance` value even if the `_geo` field is not displayed - [x] Store the HNSW on possibly multiple LMDB values. - [ ] Measure it and make it faster if needed - [ ] Export the `ReadableSlices` type into a small external crate - [x] Accept an `_vectors` field instead of the `_vector` one. - [x] Normalize all vectors. - [ ] Remove the `_vectors` field from the default searchable attributes (as we do with `_geo`?). - [ ] Correctly compute the candidates by remembering the documents having a valid `_vectors` field. - [ ] Return the right errors: - [ ] Return an error when the query vector is not the same length as the vectors in the HNSW. - [ ] We must return the user document id that triggered the vector dimension issue. - [x] If an indexation error occurs. - [ ] Fix the error codes when using the search route. - [ ] ~~Introduce some settings:~~ We currently ensure that the vector length is consistent over the whole set of documents and return an error for when a vector dimension doesn't follow the current number of dimensions. - [ ] The length of the vector the user will provide. - [ ] The distance function (we only support dot as of now). - [ ] Introduce other distance functions - [ ] Euclidean - [ ] Dot Product - [ ] Cosine - [ ] Make them SIMD optimized - [ ] Give credit to qdrant - [ ] Add tests. - [ ] Write a mini spec. - [ ] Release it in v1.3 as an experimental feature. Co-authored-by: Clément Renault <clement@meilisearch.com> Co-authored-by: Kerollmops <clement@meilisearch.com>	2023-06-27 11:17:07 +00:00
Clément Renault	e69be93e42	Log warn about using both q and vector field parameters	2023-06-27 12:32:44 +02:00
Clément Renault	b2b413db12	Return all the _semanticScore values in the documents	2023-06-27 12:32:43 +02:00
Clément Renault	30741d17fa	Change the TODO message	2023-06-27 12:32:43 +02:00
Clément Renault	ebad1f396f	Remove the useless euclidean distance implementation	2023-06-27 12:32:43 +02:00
Clément Renault	29d8268c94	Fix the vector query part by using the correct universe	2023-06-27 12:32:43 +02:00
Clément Renault	63bfe1cee2	Ignore when there are too many vectors	2023-06-27 12:32:43 +02:00
Clément Renault	f3e4d70638	Send analytics about the query vector length	2023-06-27 12:32:43 +02:00
Kerollmops	eecf20f109	Introduce a new invalid_vector_store	2023-06-27 12:32:42 +02:00
Kerollmops	816d7ed174	Update the Vector Store product feature link	2023-06-27 12:32:42 +02:00
Louis Dureuil	864ad2a23c	Check that vector store feature is enabled	2023-06-27 12:32:42 +02:00
Kerollmops	66fb5c150c	Rename _semanticSimilarity into _semanticScore	2023-06-27 12:32:42 +02:00
Kerollmops	7c2f5f77b8	Make clippy and fmt happy	2023-06-27 12:32:42 +02:00
Kerollmops	66b8cfd8c8	Introduce a way to store the HNSW on multiple LMDB entries	2023-06-27 12:32:42 +02:00
Kerollmops	ff3664431f	Make rustfmt happy	2023-06-27 12:32:42 +02:00
Kerollmops	531748c536	Return a user error when the _vectors type is invalid	2023-06-27 12:32:41 +02:00
Kerollmops	7aa1275337	Display the _semanticSimilarity even if the `_vectors` field is not displayed	2023-06-27 12:32:41 +02:00
Kerollmops	737aec1705	Expose an _semanticSimilarity as a dot product in the documents	2023-06-27 12:32:41 +02:00
Kerollmops	3e3c743392	Make Rustfmt happy	2023-06-27 12:32:41 +02:00
Kerollmops	5c5a4e075d	Make clippy happy	2023-06-27 12:32:41 +02:00
Kerollmops	ab9f2269aa	Normalize the vectors during indexation and search	2023-06-27 12:32:41 +02:00
Kerollmops	321ec5f3fa	Accept multiple vectors by documents using the _vectors field	2023-06-27 12:32:40 +02:00
Kerollmops	1b2923f7c0	Return the vector in the output of the search routes	2023-06-27 12:32:40 +02:00
Kerollmops	717d4fddd4	Remove the unused distance	2023-06-27 12:32:40 +02:00
Kerollmops	a7e0f0de89	Introduce a new error message for invalid vector dimensions	2023-06-27 12:32:40 +02:00
Kerollmops	3b560ef7d0	Make clippy happy	2023-06-27 12:32:40 +02:00
Kerollmops	2cf747cb89	Fix the tests	2023-06-27 12:32:40 +02:00
Kerollmops	3c31e1cdd1	Support more pages but in an ugly way	2023-06-27 12:32:39 +02:00
Kerollmops	23eaaf1001	Change the name of the distance module	2023-06-27 12:32:39 +02:00
Kerollmops	c2a402f3ae	Implement an ugly deletion of values in the HNSW	2023-06-27 12:32:39 +02:00
Kerollmops	436a10bef4	Replace the euclidean with a dot product	2023-06-27 12:32:39 +02:00
Kerollmops	8debf6fe81	Use a basic euclidean distance function	2023-06-27 12:32:39 +02:00
Kerollmops	c79e82c62a	Move back to the hnsw crate This reverts commit 7a4b6c065482f988b01298642f4c18775503f92f.	2023-06-27 12:32:39 +02:00
Kerollmops	aca305bb77	Log more to make sure we insert vectors in the hgg data-structure	2023-06-27 12:32:38 +02:00
Kerollmops	5816008139	Introduce an optimized version of the euclidean distance function	2023-06-27 12:32:38 +02:00
Kerollmops	268a9ef416	Move to the hgg crate	2023-06-27 12:32:38 +02:00
Clément Renault	642b0f3a1b	Expose a new vector field on the search route	2023-06-27 12:32:38 +02:00
Clément Renault	cad90e8cbc	Add a vector field to the search routes	2023-06-27 12:32:38 +02:00
Clément Renault	4571e512d2	Store the vectors in an HNSW in LMDB	2023-06-27 12:32:38 +02:00
Clément Renault	7ac2f1489d	Extract the vectors from the documents	2023-06-27 12:32:37 +02:00
Clément Renault	34349faeae	Create a new _vector extractor	2023-06-27 12:32:37 +02:00
meili-bors[bot]	ed0a5be4b6	Merge #3853 3853: docs: fixed some broken links r=gillian-meilisearch a=0xflotus Some of the links in the README file were broken. Co-authored-by: 0xflotus <0xflotus@gmail.com>	2023-06-27 10:30:13 +00:00
meili-bors[bot]	f105df6599	Merge #3850 3850: Experimental features r=Kerollmops a=dureuill # Pull Request ## Related issue - Fixes https://github.com/meilisearch/meilisearch/issues/3857 - Related to https://github.com/meilisearch/meilisearch/issues/3771 ## What does this PR do? ### Example <details> <summary>Using the feature to enable `scoreDetails`</summary> ```json ❯ curl \ -X POST 'http://localhost:7700/indexes/index-word-count-10-count/search' \ -H 'Content-Type: application/json' \ --data-binary '{ "q": "Batman", "limit": 1, "showRankingScoreDetails": true, "attributesToRetrieve": ["title"]}' \| jsonxf { "message": "Computing score details requires enabling the `score details` experimental feature. See https://github.com/meilisearch/product/discussions/674", "code": "feature_not_enabled", "type": "invalid_request", "link": "https://docs.meilisearch.com/errors#feature_not_enabled" } ``` ```json ❯ curl \ -X PATCH 'http://localhost:7700/experimental-features/' \ -H 'Content-Type: application/json' \ --data-binary '{ "scoreDetails": true }' {"scoreDetails":true,"vectorSearch":false} ``` ```json ❯ curl \ -X POST 'http://localhost:7700/indexes/index-word-count-10-count/search' \ -H 'Content-Type: application/json' \ --data-binary '{ "q": "Batman", "limit": 1, "showRankingScoreDetails": true, "attributesToRetrieve": ["title"]}' \| jsonxf { "hits": [ { "title": "Batman", "_rankingScoreDetails": { "words": { "order": 0, "matchingWords": 1, "maxMatchingWords": 1, "score": 1.0 }, "typo": { "order": 1, "typoCount": 0, "maxTypoCount": 1, "score": 1.0 }, "proximity": { "order": 2, "score": 1.0 }, "attribute": { "order": 3, "attribute_ranking_order_score": 1.0, "query_word_distance_score": 1.0, "score": 1.0 }, "exactness": { "order": 4, "matchType": "exactMatch", "score": 1.0 } } } ], "query": "Batman", "processingTimeMs": 3, "limit": 1, "offset": 0, "estimatedTotalHits": 46 } ``` </details> ### User standpoint - Add new route GET/POST/PATCH/DELETE `/experimental-features` to switch on or off some of the experimental features in a manner persistent between instance restarts - Use these new routes to allow setting on/off the following experimental features: - vector store TODO: fill in issue - score details (related to https://github.com/meilisearch/meilisearch/issues/3771) - Make the way of checking feature availability and error message uniform for the Prometheus metrics experimental feature - Save the enabled features in dump, restore from dumps - TODO: tests: - Test new security permissions (do they allow access with ALL, do they prevent access when missing) - Test dump behavior, in particular ability to import existing v6 dumps - Test basic behavior when calling the rule ### Implementation standpoint - New DB "experimental-features" - dumps are modified to save the state of that new DB as a `experimental-features.json` file, that is then loaded back when importing the dump. This doesn't change the dump version, as the file is optional and it missing will not cause the dump to fail Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-06-26 15:13:43 +00:00
Louis Dureuil	13e9b4c2e5	Add dump support	2023-06-26 16:29:43 +02:00
Louis Dureuil	5a83cecb0f	fix tests	2023-06-26 16:29:43 +02:00
Louis Dureuil	cca6e47ec1	Errors when GETting metrics without the feature gate	2023-06-26 16:29:43 +02:00
Louis Dureuil	6196a53668	Gate score_details behind a runtime experimental feature flag	2023-06-26 16:29:43 +02:00
Louis Dureuil	bb6448dc2e	Compute instance features from CLI options	2023-06-26 16:29:43 +02:00
Louis Dureuil	eef9293630	New route to set some experimental features	2023-06-26 16:29:43 +02:00
Louis Dureuil	dac77dfd14	Add new permissions for experimental-features route	2023-06-26 16:29:43 +02:00
Louis Dureuil	072d81843f	Persistently save to DB the status of experimental features	2023-06-26 16:29:43 +02:00
Louis Dureuil	29ec02d4d4	Add meilisearch_types::features module	2023-06-26 16:09:03 +02:00
ManyTheFish	9d2a12821d	Use insta snapshot	2023-06-26 14:56:19 +02:00
ManyTheFish	63ca25290b	Take into account small Review requests	2023-06-26 14:56:19 +02:00
ManyTheFish	59f64a5256	Return an error when an attribute is not searchable	2023-06-26 14:56:19 +02:00
ManyTheFish	dc391deca0	Reverse assert comparison to have a consistent error message	2023-06-26 14:55:57 +02:00
ManyTheFish	114f878205	Rename restrictSearchableAttributes into attributesToSearchOn	2023-06-26 14:55:57 +02:00
ManyTheFish	42709ea9a5	Fix clippy warnings	2023-06-26 14:55:57 +02:00
ManyTheFish	993b0d012c	Remove proximity_ranking_rule_order test, fixing this test would force us to create a fid_word_pair_proximity_docids and a fid_word_prefix_pair_proximity_docids databases which may multiply the keys of word_pair_proximity_docids and word_prefix_pair_proximity_docids by the number of attributes in the searchable_attributes list	2023-06-26 14:55:57 +02:00
ManyTheFish	fb8fa07169	Restrict field ids in search context	2023-06-26 14:55:57 +02:00
ManyTheFish	0ccf1e2e40	Allow the search cache to store owned values	2023-06-26 14:55:57 +02:00
ManyTheFish	9680e1e41f	Introduce a BytesDecodeOwned trait in heed_codecs	2023-06-26 14:55:14 +02:00
ManyTheFish	a61ca4066e	Add tests	2023-06-26 14:55:14 +02:00
ManyTheFish	461b5118bd	Add API search setting	2023-06-26 14:55:14 +02:00
Tamo	a3716c5678	add the new parameter to the search builder of milli	2023-06-26 14:55:14 +02:00
meili-bors[bot]	2d34005965	Merge #3821 3821: Add normalized and detailed scores to documents returned by a query r=dureuill a=dureuill # Pull Request ## Related issue Fixes #3771 ## What does this PR do? ### User standpoint <details> <summary>Request ranking score</summary> ``` echo '{ "q": "Badman dark knight returns", "showRankingScore": true, "limit": 10, "attributesToRetrieve": ["title"] }' \| mieli search -i index-word-count-10-count ``` </details> <details> <summary>Response</summary> ```json { "hits": [ { "title": "Batman: The Dark Knight Returns, Part 1", "_rankingScore": 0.947520325203252 }, { "title": "Batman: The Dark Knight Returns, Part 2", "_rankingScore": 0.947520325203252 }, { "title": "Batman Unmasked: The Psychology of the Dark Knight", "_rankingScore": 0.6657594086021505 }, { "title": "Legends of the Dark Knight: The History of Batman", "_rankingScore": 0.6654905913978495 }, { "title": "Angel and the Badman", "_rankingScore": 0.2196969696969697 }, { "title": "Angel and the Badman", "_rankingScore": 0.2196969696969697 }, { "title": "Batman", "_rankingScore": 0.11553030303030302 }, { "title": "Batman Begins", "_rankingScore": 0.11553030303030302 }, { "title": "Batman Returns", "_rankingScore": 0.11553030303030302 }, { "title": "Batman Forever", "_rankingScore": 0.11553030303030302 } ], "query": "Badman dark knight returns", "processingTimeMs": 12, "limit": 10, "offset": 0, "estimatedTotalHits": 46 } ``` </details> - If adding a `showRankingScore` parameter to the search query, then documents returned by a search now contain an additional field `_rankingScore` that is a float bigger than 0 and lower or equal to 1.0. This field represents the relevancy of the document, relatively to the search query and the settings of the index, with 1.0 meaning "perfect match" and 0 meaning "not matching the query" (Meilisearch should never return documents not matching the query at all). - The `sort` and `geosort` ranking rules do not influence the `_rankingScore`. <details> <summary>Request detailed ranking scores</summary> ``` echo '{ "q": "Badman dark knight returns", "showRankingScoreDetails": true, "limit": 5, "attributesToRetrieve": ["title"] }' \| mieli search -i index-word-count-10-count ``` </details> <details> <summary>Response</summary> ```json { "hits": [ { "title": "Batman: The Dark Knight Returns, Part 1", "_rankingScoreDetails": { "words": { "order": 0, "matchingWords": 4, "maxMatchingWords": 4, "score": 1.0 }, "typo": { "order": 1, "typoCount": 1, "maxTypoCount": 4, "score": 0.8 }, "proximity": { "order": 2, "score": 0.9545454545454546 }, "attribute": { "order": 3, "attributes_ranking_order": 1.0, "attributes_query_word_order": 0.926829268292683, "score": 0.926829268292683 }, "exactness": { "order": 4, "matchType": "noExactMatch", "score": 0.26666666666666666 } } }, { "title": "Batman: The Dark Knight Returns, Part 2", "_rankingScoreDetails": { "words": { "order": 0, "matchingWords": 4, "maxMatchingWords": 4, "score": 1.0 }, "typo": { "order": 1, "typoCount": 1, "maxTypoCount": 4, "score": 0.8 }, "proximity": { "order": 2, "score": 0.9545454545454546 }, "attribute": { "order": 3, "attributes_ranking_order": 1.0, "attributes_query_word_order": 0.926829268292683, "score": 0.926829268292683 }, "exactness": { "order": 4, "matchType": "noExactMatch", "score": 0.26666666666666666 } } }, { "title": "Batman Unmasked: The Psychology of the Dark Knight", "_rankingScoreDetails": { "words": { "order": 0, "matchingWords": 3, "maxMatchingWords": 4, "score": 0.75 }, "typo": { "order": 1, "typoCount": 1, "maxTypoCount": 3, "score": 0.75 }, "proximity": { "order": 2, "score": 0.6666666666666666 }, "attribute": { "order": 3, "attributes_ranking_order": 1.0, "attributes_query_word_order": 0.8064516129032258, "score": 0.8064516129032258 }, "exactness": { "order": 4, "matchType": "noExactMatch", "score": 0.25 } } }, { "title": "Legends of the Dark Knight: The History of Batman", "_rankingScoreDetails": { "words": { "order": 0, "matchingWords": 3, "maxMatchingWords": 4, "score": 0.75 }, "typo": { "order": 1, "typoCount": 1, "maxTypoCount": 3, "score": 0.75 }, "proximity": { "order": 2, "score": 0.6666666666666666 }, "attribute": { "order": 3, "attributes_ranking_order": 1.0, "attributes_query_word_order": 0.7419354838709677, "score": 0.7419354838709677 }, "exactness": { "order": 4, "matchType": "noExactMatch", "score": 0.25 } } }, { "title": "Angel and the Badman", "_rankingScoreDetails": { "words": { "order": 0, "matchingWords": 1, "maxMatchingWords": 4, "score": 0.25 }, "typo": { "order": 1, "typoCount": 0, "maxTypoCount": 1, "score": 1.0 }, "proximity": { "order": 2, "score": 1.0 }, "attribute": { "order": 3, "attributes_ranking_order": 1.0, "attributes_query_word_order": 0.8181818181818182, "score": 0.8181818181818182 }, "exactness": { "order": 4, "matchType": "noExactMatch", "score": 0.3333333333333333 } } } ], "query": "Badman dark knight returns", "processingTimeMs": 9, "limit": 5, "offset": 0, "estimatedTotalHits": 46 } ``` </details> - If adding a `showRankingScoreDetails` parameter to the search query, then the returned documents will now contain an additional `_rankingScoreDetails` field that is a JSON object containing one field per ranking rule that was applied, whose value is a JSON object with the following fields: - `order`: a number indicating the order this rule was applied (0 is the first applied ranking rule) - `score` (except for `sort` and `geosort`): a float indicating how the document matched this particular rule. - other fields that are specific to the rule, indicating for example how many words matched for a document and how many typos were counted in a matching document. - If the `displayableAttributes` list is defined in the settings of the index, any ranking rule using an attribute not part of that list will be marked as `<hidden-rule>` in the `_rankingScoreDetails`. - Search queries that are part of a `multi-search` requests are modified in the same way and each of the queries can take the `showRankingScore` and `showRankingScoreDetails` parameters independently. The results are still returned in separate lists and providing a unified list of results between multiple queries is not in the scope of this PR (but is unblocked by this PR and can be done manually by using the scores of the various documents). ### Implementation standpoint - Fix difference in how the position of terms were computed at indexing time and query time: this difference meant that a query containing a hard separator would fail the exactness check. - Fix the id reported by the sort ranking rule (very minor) - Change how the cost of removing words is computed. After this change the cost no longer works for any other ranking rule than `words`. Also made `words` have a cost of 0 such that the entire cost of `words` is given by the termRemovalStrategy. The new cost computation makes it so the score is computed in a way consistent with the number of words in the query. Additionally, the words that appear in phrases in the query are also counted as matching words. - When any score computation is requested through `showRankingScore` or `showRankingScoreDetails`, remove optimization where ranking rules are not executed on buckets of a single document: this is important to allow the computation of an accurate score. - add virtual conditions to fid and position to always have the max cost: this ensures that the score is independent from the dataset - the Position ranking rule now takes into account the distance to the position of the word in the query instead of the distance to the position 0. - modified proximity ranking rule cost calculation so that the cost is 0 for documents that are perfectly matching the query - Add a new `milli::score_details` module containing all the types that are involved in score computation. - Make it so a bucket of result now contains a `ScoreDetails` and changed the ranking rules to produce their `ScoreDetails`. - Expose the scores in the REST API. - Add very light analytics for scoring. - Update the search tests to add the expected scores. Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-06-26 09:32:43 +00:00
Louis Dureuil	62eefcda6e	Change and add links to the Cloud	2023-06-26 09:17:15 +02:00
0xflotus	85a24775c5	Update README.md	2023-06-23 12:25:53 +02:00
0xflotus	6b0e9b9a7f	Update README.md	2023-06-23 12:20:43 +02:00
0xflotus	b18c57ea7f	docs: fixed some broken links Some of the links in the README file were broken.	2023-06-23 12:18:43 +02:00
Cong Chen	6d4981ec25	Expose lastUpdate and isIndexing in /stats endpoint	2023-06-23 07:24:25 +08:00
meili-bors[bot]	040b5a5b6f	Merge #3842 3842: fix some typos r=dureuill a=cuishuang # Pull Request ## Related issue Fixes #<issue_number> ## What does this PR do? - fix some typos ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: cui fliter <imcusg@gmail.com>	2023-06-22 18:01:10 +00:00
cui fliter	530a3e2df3	fix some typos Signed-off-by: cui fliter <imcusg@gmail.com>	2023-06-22 21:59:00 +08:00
Louis Dureuil	11d32ad192	Add very light analytics for scoring	2023-06-22 12:39:14 +02:00
Louis Dureuil	d26e9a96ec	Add score details to new search tests	2023-06-22 12:39:14 +02:00
Louis Dureuil	49c8bc4de6	Fix tests	2023-06-22 12:39:14 +02:00
Louis Dureuil	da833eb095	Expose the scores and detailed scores in the API	2023-06-22 12:39:14 +02:00
Louis Dureuil	701d44bd91	Store the scores for each bucket Remove optimization where ranking rules are not executed on buckets of a single document when the score needs to be computed	2023-06-22 12:39:14 +02:00
Louis Dureuil	c621a250a7	Score for graph based ranking rules Count phrases in matchingWords and maxMatchingWords	2023-06-22 12:39:14 +02:00
Louis Dureuil	8939e85f60	Add rank_to_score for graph based ranking rules	2023-06-22 12:39:14 +02:00
Louis Dureuil	fa41d2489e	Score for sort	2023-06-22 12:39:14 +02:00
Louis Dureuil	59c5b992c2	Score for geosort	2023-06-22 12:39:14 +02:00
Louis Dureuil	2ea8194c18	Score for exact_attributes	2023-06-22 12:39:14 +02:00
Louis Dureuil	421df64602	RankingRuleOutput now contains a Score	2023-06-22 12:39:14 +02:00
Louis Dureuil	c0fca6f884	Add score_details	2023-06-22 12:39:14 +02:00
Filip Bachul	9015a8e8d9	Merge branch 'main' into cymruu/payload-unit-test	2023-06-21 09:26:50 +02:00
meili-bors[bot]	28404d56b7	Merge #3799 3799: Fix error messages in `check-release.sh` r=curquiza a=vvv - `check_tag`: Report file name correctly. Use named variables. - Introduce `read_version` helper function. Simplify the implementation. - Show meaningful error message if `GITHUB_REF` is not set or its format is incorrect. Co-authored-by: Valeriy V. Vorotyntsev <valery.vv@gmail.com>	2023-06-20 13:35:33 +00:00
meili-bors[bot]	262c1f2baf	Merge #3844 3844: Fix SDK CI (again) r=curquiza a=curquiza Following this PR: https://github.com/meilisearch/meilisearch/pull/3813 Sorry `@Kerollmops,` here is (I hope) the latest fix 🙏 I made tests last time that were not sufficient. I really did a lot this time. I hope I have not missed anything. Co-authored-by: curquiza <clementine@meilisearch.com>	2023-06-20 13:01:07 +00:00
Valeriy V. Vorotyntsev	cfed349aa3	Fix error messages in `check-release.sh` - `check_tag`: Report file name correctly. Use named variables. - Introduce `read_version` helper function. Simplify the implementation. - Show meaningful error message if `GITHUB_REF` is not set or its format is incorrect.	2023-06-20 13:58:09 +03:00
Louis Dureuil	f050634b1e	add virtual conditions to fid and position to always have the max cost	2023-06-20 10:07:18 +02:00
Louis Dureuil	becf1f066a	Change how the cost of removing words is computed	2023-06-20 09:45:43 +02:00
Louis Dureuil	701d299369	Remove out-of-date comment	2023-06-20 09:45:42 +02:00
Louis Dureuil	a20e4d447c	Position now takes into account the distance to the position of the word in the query it used to be based on the distance to the position 0	2023-06-20 09:45:42 +02:00
Louis Dureuil	af57c3c577	Proximity costs 0 for documents that are perfectly matching	2023-06-20 09:45:42 +02:00
Louis Dureuil	0c40ef6911	Fix sort id	2023-06-20 09:45:42 +02:00
curquiza	bbc9f68ff5	Use the input from the previous job instead of the workflow dispatch	2023-06-19 18:49:15 +02:00
meili-bors[bot]	45636d315c	Merge #3670 3670: Fix addition deletion bug r=irevoire a=irevoire The first commit of this PR is a revert of https://github.com/meilisearch/meilisearch/pull/3667. It re-enable the auto-batching of addition and deletion of tasks. No new changes have been introduced outside of `milli`. So all the changes you see on the autobatcher have actually already been reviewed. It fixes https://github.com/meilisearch/meilisearch/issues/3440. ### What was happening? The issue was that the `external_documents_ids` generated in the `transform` were used in a very strange way that wasn’t compatible with the deletion of documents. Instead of doing a clear merge between the external document IDs of the DB and the one returned by the transform + writing it on disk, we were doing some weird tricks with the soft-deleted to avoid writing the fst on disk as much as possible. The new algorithm may be a bit slower but is way more straightforward and doesn’t change depending on if the soft deletion was used or not. Here is a list of the changes introduced: 1. We now do a clear distinction between the `new_external_documents_ids` coming from the transform and only held on RAM and the `external_documents_ids` coming from the DB. 2. The `new_external_documents_ids` (coming out of the transform) are now represented as an `fst`. We don't need to struggle with the hard, soft distinction + the soft_deleted => That's easier to understand 3. When indexing documents, we merge the `external_documents_ids` coming from the DB and the `new_external_documents_ids` coming from the transform. ### Other things introduced in this PR Since we constantly have to write small, very specialized fuzzers for this kind of bug, we decided to push the one used to reproduce this bug. It's not perfect, but it's easy to improve in the future. It'll also run for as long as possible on every merge on the main branch. Co-authored-by: Tamo <tamo@meilisearch.com> Co-authored-by: Loïc Lecrenier <loic.lecrenier@icloud.com>	2023-06-19 09:09:30 +00:00
meili-bors[bot]	cb9d78fc7f	Merge #3835 3835: Add more documentation to graph-based ranking rule algorithms + comment cleanup r=Kerollmops a=loiclec In addition to documenting the `cheapest_path.rs` file, this PR cleans up a few outdated comments as well as some TODOs. These TODOs have been moved to https://github.com/meilisearch/meilisearch/issues/3776 Co-authored-by: Loïc Lecrenier <loic.lecrenier@icloud.com>	2023-06-15 15:30:24 +00:00
meili-bors[bot]	01d2ee5cc1	Merge #3836 3836: Remove trailing whitespace in snapshots r=dureuill a=dureuill # Pull Request ## Related issue No issue, maintenance ## What does this PR do? - Remove trailing whitespace in snapshots by adding a trailing `\|` at the end of lines that would previously end with fixed-width integers - This allows contributors whose editor is configured to remove trailing whitespace not to modify the tests when changing an unrelated part of the file containing the tests Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-06-14 13:00:52 +00:00
Louis Dureuil	e0c4682758	Fix tests	2023-06-14 13:30:52 +02:00
Louis Dureuil	d9b4b39922	Add trailing pipe to the snapshots so it doesn't end with trailing whitespace	2023-06-14 13:30:52 +02:00
Loïc Lecrenier	2da86b31a6	Remove comments and add documentation	2023-06-14 12:39:42 +02:00
Loïc Lecrenier	4e81445d42	Stop the fuzzer after an hour	2023-06-12 15:30:51 +02:00
meili-bors[bot]	4829348d6e	Merge #3813 3813: Fix SDK CI for scheduled jobs r=curquiza a=curquiza The SDK CI does not run for the scheduled job (`cron`) every day, and only works for manual triggers. I added a job to define the Docker image we use depending on the event: `worflow_dispatch` = manual triggering, or `scheduled` = cron jobs Co-authored-by: curquiza <clementine@meilisearch.com>	2023-06-12 08:41:03 +00:00
meili-bors[bot]	047d22fcb1	Merge #3824 3824: Changes the way words are counted in the word count DB r=ManyTheFish a=dureuill # Pull Request ## Related issue Fixes https://github.com/meilisearch/meilisearch/issues/3823 ## What does this PR do? - Apply offset when parsing query that is consistent with the indexing ### DB breaking changes - Count the number of words in `field_id_word_count_docids` - raise limit of word count for storing the entry in the DB from 10 to 30 Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-06-08 13:26:05 +00:00
Louis Dureuil	a2a3b8c973	Fix offset difference between query and indexing for hard separators	2023-06-08 12:07:12 +02:00
Louis Dureuil	9f37b61666	DB BREAKING: raise limit of word count from 10 to 30.	2023-06-08 12:07:12 +02:00
Louis Dureuil	c15c076da9	DB BREAKING: Count the number of words in field_id_word_count_docids	2023-06-08 12:07:11 +02:00
meili-bors[bot]	9dcf1da59d	Merge #3819 3819: Remove the `docid_word_positions` database r=Kerollmops a=loiclec Remove the `docid_word_positions` database, which was only used during deletion operations. In the process, also fixes https://github.com/meilisearch/meilisearch/issues/3816 Co-authored-by: Loïc Lecrenier <loic.lecrenier@icloud.com>	2023-06-07 09:53:25 +00:00
Loïc Lecrenier	8628a0c856	Remove docid_word_positions_db + fix deletion bug That would happen when a word was deleted from all exact attributes but not all regular attributes.	2023-06-07 10:52:50 +02:00
meili-bors[bot]	c1e3cc04b0	Merge #3811 3811: Bring back changes from `release-v1.2.0` to `main` r=Kerollmops a=curquiza Co-authored-by: Loïc Lecrenier <loic.lecrenier@me.com> Co-authored-by: meili-bors[bot] <89034592+meili-bors[bot]@users.noreply.github.com> Co-authored-by: Tamo <tamo@meilisearch.com> Co-authored-by: Filip Bachul <filipbachul@gmail.com> Co-authored-by: Kerollmops <clement@meilisearch.com> Co-authored-by: ManyTheFish <many@meilisearch.com> Co-authored-by: Clément Renault <clement@meilisearch.com>	2023-06-06 13:10:24 +00:00
meili-bors[bot]	d96d8bb0dd	Merge #3789 3789: Improve the metrics r=dureuill a=irevoire # Pull Request ## Related issue Implements https://github.com/meilisearch/meilisearch/issues/3790 Associated specification: https://github.com/meilisearch/specifications/pull/242 ## Be cautious; it's DB-breaking 😱 While reviewing and after merging this PR, be cautious; if you already have a `data.ms` and run meilisearch with this code on it, it won't work because we need to cache a new information on the index stats (that are backed up on disk). You'll get internal errors. ### About the breaking-change label We only break the API of the metrics route, which does not pose any problem since it's experimental. ## What does this PR do? - Create a method to get the « facet distribution » of the task queue. - Prefix all the metrics by `meilisearch_` - Add the real database size used by meilisearch - Add metrics on the task queue - Update the grafana dashboard to these new changes - Move the dashboard to the `assets` directory - Provide a new prometheus file to scrape meilisearch easily Co-authored-by: Tamo <tamo@meilisearch.com>	2023-06-06 11:44:54 +00:00
Tamo	4a3405afec	comment the stats method	2023-06-06 12:59:58 +02:00
Tamo	3cfd653db1	Apply suggestions from code review Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-06-06 11:38:41 +02:00
curquiza	b6b6a80b76	Fix SDK CI for scheduled jobs	2023-06-06 10:38:05 +02:00
Clémentine U. - curqui	f3e2f79290	Merge branch 'main' into tmp-release-v1.2.0	2023-06-05 18:36:28 +02:00
meili-bors[bot]	f517274d1f	Merge #3788 3788: Use `RoaringBitmap::deserialize_unchecked_from` to reduce the deserialization time r=irevoire a=Kerollmops This pull request replaces the `RoaringBitmap::deserialize_from` methods with the `deserialize_unchecked_from` to avoid doing too much checks. We know the written bitmaps are valid as we do not disable the checks during the indexation phase. I did a small test with #3780 and discovered that the deserialization time changed from 32% to 9.46% when using these changes. It seems it was low-hanging fruit hidden behind a leaf. Co-authored-by: Kerollmops <clement@meilisearch.com>	2023-06-05 09:20:30 +00:00
meili-bors[bot]	3f41bc642a	Merge #3804 #3805 3804: Bump svenstaro/upload-release-action from 2.5.0 to 2.6.1 r=curquiza a=dependabot[bot] Bumps [svenstaro/upload-release-action](https://github.com/svenstaro/upload-release-action) from 2.5.0 to 2.6.1. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/svenstaro/upload-release-action/releases">svenstaro/upload-release-action's releases</a>.</em></p> <blockquote> <h2>2.6.1</h2> <ul> <li>Do not overwrite body or name if empty <a href="https://redirect.github.com/svenstaro/upload-release-action/pull/108">#108</a> (thanks <a href="https://github.com/regevbr"><code>`@regevbr</code></a>)</li>` </ul> <h2>2.6.0</h2> <ul> <li>Add <code>make_latest</code> input parameter. Can be set to <code>false</code> to prevent the created release from being marked as the latest release for the repository <a href="https://redirect.github.com/svenstaro/upload-release-action/pull/100">#100</a> (thanks <a href="https://github.com/brandonkelly"><code>`@brandonkelly</code></a>)</li>` <li>Don't try to upload empty files <a href="https://redirect.github.com/svenstaro/upload-release-action/pull/102">#102</a> (thanks <a href="https://github.com/Loyalsoldier"><code>`@Loyalsoldier</code></a>)</li>` <li>Bump all deps <a href="https://redirect.github.com/svenstaro/upload-release-action/pull/105">#105</a></li> <li><code>overwrite</code> option also overwrites name and body <a href="https://redirect.github.com/svenstaro/upload-release-action/pull/106">#106</a> (thanks <a href="https://github.com/regevbr"><code>`@regevbr</code></a>)</li>` <li>Add <code>promote</code> option to allow prereleases to be promoted <a href="https://redirect.github.com/svenstaro/upload-release-action/pull/74">#74</a> (thanks <a href="https://github.com/regevbr"><code>`@regevbr</code></a>)</li>` </ul> </blockquote> </details> <details> <summary>Changelog</summary> <p><em>Sourced from <a href="https://github.com/svenstaro/upload-release-action/blob/master/CHANGELOG.md">svenstaro/upload-release-action's changelog</a>.</em></p> <blockquote> <h2>[2.6.1] - 2023-05-31</h2> <ul> <li>Do not overwrite body or name if empty <a href="https://redirect.github.com/svenstaro/upload-release-action/pull/108">#108</a> (thanks <a href="https://github.com/regevbr"><code>`@regevbr</code></a>)</li>` </ul> <h2>[2.6.0] - 2023-05-23</h2> <ul> <li>Add <code>make_latest</code> input parameter. Can be set to <code>false</code> to prevent the created release from being marked as the latest release for the repository <a href="https://redirect.github.com/svenstaro/upload-release-action/pull/100">#100</a> (thanks <a href="https://github.com/brandonkelly"><code>`@brandonkelly</code></a>)</li>` <li>Don't try to upload empty files <a href="https://redirect.github.com/svenstaro/upload-release-action/pull/102">#102</a> (thanks <a href="https://github.com/Loyalsoldier"><code>`@Loyalsoldier</code></a>)</li>` <li>Bump all deps <a href="https://redirect.github.com/svenstaro/upload-release-action/pull/105">#105</a></li> <li><code>overwrite</code> option also overwrites name and body <a href="https://redirect.github.com/svenstaro/upload-release-action/pull/106">#106</a> (thanks <a href="https://github.com/regevbr"><code>`@regevbr</code></a>)</li>` <li>Add <code>promote</code> option to allow prereleases to be promoted <a href="https://redirect.github.com/svenstaro/upload-release-action/pull/74">#74</a> (thanks <a href="https://github.com/regevbr"><code>`@regevbr</code></a>)</li>` </ul> </blockquote> </details> <details> <summary>Commits</summary> <ul> <li><a href="`2b9d2847a9`"><code>2b9d284</code></a> 2.6.1</li> <li><a href="`f9beb0ad08`"><code>f9beb0a</code></a> Merge pull request <a href="https://redirect.github.com/svenstaro/upload-release-action/issues/108">#108</a> from regevbr/<a href="https://redirect.github.com/svenstaro/upload-release-action/issues/107">#107</a></li> <li><a href="`1662cfa449`"><code>1662cfa</code></a> fix <a href="https://redirect.github.com/svenstaro/upload-release-action/issues/197">#197</a> - do not overwrite, if empty</li> <li><a href="`a5002416a0`"><code>a500241</code></a> Document running npm update after changing version</li> <li><a href="`58d5258088`"><code>58d5258</code></a> 2.6.0</li> <li><a href="`ffc1afa9c0`"><code>ffc1afa</code></a> Update CHANGELOG</li> <li><a href="`24bced81d9`"><code>24bced8</code></a> Merge pull request <a href="https://redirect.github.com/svenstaro/upload-release-action/issues/74">#74</a> from regevbr/body</li> <li><a href="`794b3152e1`"><code>794b315</code></a> fix <a href="https://redirect.github.com/svenstaro/upload-release-action/issues/42">#42</a> - overwrite body and name as well</li> <li><a href="`b00963776a`"><code>b009637</code></a> fix <a href="https://redirect.github.com/svenstaro/upload-release-action/issues/42">#42</a> - overwrite body and name as well</li> <li><a href="`210500d479`"><code>210500d</code></a> fix <a href="https://redirect.github.com/svenstaro/upload-release-action/issues/42">#42</a> - overwrite body and name as well</li> <li>Additional commits viewable in <a href="https://github.com/svenstaro/upload-release-action/compare/2.5.0...2.6.1">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=svenstaro/upload-release-action&package-manager=github_actions&previous-version=2.5.0&new-version=2.6.1)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) You can trigger a rebase of this PR by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) </details> 3805: Bump actions/setup-go from 3 to 4 r=curquiza a=dependabot[bot] Bumps [actions/setup-go](https://github.com/actions/setup-go) from 3 to 4. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/actions/setup-go/releases">actions/setup-go's releases</a>.</em></p> <blockquote> <h2>v4.0.0</h2> <p>In scope of release we enable cache by default. The action won’t throw an error if the cache can’t be restored or saved. The action will throw a warning message but it won’t stop a build process. The cache can be disabled by specifying <code>cache: false</code>.</p> <pre lang="yaml"><code>steps: - uses: actions/checkout@v3 - uses: actions/setup-go@v4 with: go-version: ‘1.19’ - run: go run hello.go </code></pre> <p>Besides, we introduce such changes as</p> <ul> <li><a href="https://redirect.github.com/actions/setup-go/pull/305">Allow to use only GOCACHE for cache</a></li> <li><a href="https://redirect.github.com/actions/setup-go/pull/315">Bump json5 from 2.2.1 to 2.2.3</a></li> <li><a href="https://redirect.github.com/actions/setup-go/pull/323">Use proper version for primary key in cache</a></li> <li><a href="https://redirect.github.com/actions/setup-go/pull/351">Always add Go bin to the PATH</a></li> <li><a href="https://redirect.github.com/actions/setup-go/pull/350">Add step warning if go-version input is empty</a></li> </ul> <h2>Add support for stable and oldstable aliases</h2> <p>In scope of this release we introduce aliases for the <code>go-version</code> input. The <code>stable</code> alias instals the latest stable version of Go. The <code>oldstable</code> alias installs previous latest minor release (the stable is 1.19.x -> the oldstable is 1.18.x).</p> <h3>Stable</h3> <pre lang="yaml"><code>steps: - uses: actions/checkout@v3 - uses: actions/setup-go@v3 with: go-version: 'stable' - run: go run hello.go </code></pre> <h3>OldStable</h3> <pre lang="yaml"><code>steps: - uses: actions/checkout@v3 - uses: actions/setup-go@v3 with: go-version: 'oldstable' - run: go run hello.go </code></pre> <h2>Add support for go.work and pass the token input through on GHES</h2> <p>In scope of this release we added <a href="https://redirect.github.com/actions/setup-go/pull/283">support for go.work file to pass it in go-version-file input</a>.</p> <pre lang="yaml"><code>steps: - uses: actions/checkout@v3 - uses: actions/setup-go@v3 </tr></table> </code></pre> </blockquote> <p>... (truncated)</p> </details> <details> <summary>Commits</summary> <ul> <li><a href="`fac708d667`"><code>fac708d</code></a> Bump <code>`@actions/cache</code>` dependency to v3.2.1 (<a href="https://redirect.github.com/actions/setup-go/issues/374">#374</a>)</li> <li><a href="`dd84a9531a`"><code>dd84a95</code></a> Update xml2js (<a href="https://redirect.github.com/actions/setup-go/issues/370">#370</a>)</li> <li><a href="`41c2024c46`"><code>41c2024</code></a> Fix glob bug in package.json scripts section (<a href="https://redirect.github.com/actions/setup-go/issues/359">#359</a>)</li> <li><a href="`8dbf352f06`"><code>8dbf352</code></a> update README fo v4 (<a href="https://redirect.github.com/actions/setup-go/issues/354">#354</a>)</li> <li><a href="`4d34df0c23`"><code>4d34df0</code></a> Update configuration files (<a href="https://redirect.github.com/actions/setup-go/issues/348">#348</a>)</li> <li><a href="`fdc0d672a1`"><code>fdc0d67</code></a> Add Go bin if go-version input is empty (<a href="https://redirect.github.com/actions/setup-go/issues/351">#351</a>)</li> <li><a href="`ebfdf6ac95`"><code>ebfdf6a</code></a> add warning if go-version is empty (<a href="https://redirect.github.com/actions/setup-go/issues/350">#350</a>)</li> <li><a href="`b27d76912e`"><code>b27d769</code></a> fix lockfileVersion (<a href="https://redirect.github.com/actions/setup-go/issues/349">#349</a>)</li> <li><a href="`c51a720768`"><code>c51a720</code></a> Enable caching by default with default input (<a href="https://redirect.github.com/actions/setup-go/issues/332">#332</a>)</li> <li><a href="`6b848af622`"><code>6b848af</code></a> Merge pull request <a href="https://redirect.github.com/actions/setup-go/issues/343">#343</a> from akv-platform/reusable-workflow</li> <li>Additional commits viewable in <a href="https://github.com/actions/setup-go/compare/v3...v4">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=actions/setup-go&package-manager=github_actions&previous-version=3&new-version=4)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) You can trigger a rebase of this PR by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) </details> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2023-06-05 08:36:22 +00:00
meili-bors[bot]	672abdb341	Merge #3803 3803: Bump Swatinem/rust-cache from 2.2.1 to 2.4.0 r=curquiza a=dependabot[bot] Bumps [Swatinem/rust-cache](https://github.com/Swatinem/rust-cache) from 2.2.1 to 2.4.0. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/Swatinem/rust-cache/releases">Swatinem/rust-cache's releases</a>.</em></p> <blockquote> <h2>v2.4.0</h2> <ul> <li>Fix cache key stability.</li> <li>Use 8 character hash components to reduce the key length, making it more readable.</li> </ul> <h2>v2.3.0</h2> <ul> <li>Add <code>cache-all-crates</code> option, which enables caching of crates installed by workflows.</li> <li>Add installed packages to cache key, so changes to workflows that install rust tools are detected and cached properly.</li> <li>Fix cache restore failures due to upstream bug.</li> <li>Fix <code>EISDIR</code> error due to globed directories.</li> <li>Update runtime <code>`@actions/cache</code>,` <code>`@actions/io</code>` and dev <code>typescript</code> dependencies.</li> <li>Update <code>npm run prepare</code> so it creates distribution files with the right line endings.</li> </ul> </blockquote> </details> <details> <summary>Changelog</summary> <p><em>Sourced from <a href="https://github.com/Swatinem/rust-cache/blob/master/CHANGELOG.md">Swatinem/rust-cache's changelog</a>.</em></p> <blockquote> <h2>2.4.0</h2> <ul> <li>Fix cache key stability.</li> <li>Use 8 character hash components to reduce the key length, making it more readable.</li> </ul> <h2>2.3.0</h2> <ul> <li>Add <code>cache-all-crates</code> option, which enables caching of crates installed by workflows.</li> <li>Add installed packages to cache key, so changes to workflows that install rust tools are detected and cached properly.</li> <li>Fix cache restore failures due to upstream bug.</li> <li>Fix <code>EISDIR</code> error due to globed directories.</li> <li>Update runtime <code>`@actions/cache</code>,` <code>`@actions/io</code>` and dev <code>typescript</code> dependencies.</li> <li>Update <code>npm run prepare</code> so it creates distribution files with the right line endings.</li> </ul> </blockquote> </details> <details> <summary>Commits</summary> <ul> <li><a href="`988c164c3d`"><code>988c164</code></a> 2.4.0</li> <li><a href="`bb80d0f127`"><code>bb80d0f</code></a> chore: use 8 character hash components (<a href="https://redirect.github.com/Swatinem/rust-cache/issues/143">#143</a>)</li> <li><a href="`ad97570a01`"><code>ad97570</code></a> fix: cache key stability (<a href="https://redirect.github.com/Swatinem/rust-cache/issues/142">#142</a>)</li> <li><a href="`060bda31e0`"><code>060bda3</code></a> 2.3.0</li> <li><a href="`865fd1f6db`"><code>865fd1f</code></a> "update dependencies and changelog"</li> <li><a href="`7c7e41ab01`"><code>7c7e41a</code></a> chore: changelog v2.3.0 (<a href="https://redirect.github.com/Swatinem/rust-cache/issues/139">#139</a>)</li> <li><a href="`68aeeba167`"><code>68aeeba</code></a> chore: use linefix to ensure platform line endings (<a href="https://redirect.github.com/Swatinem/rust-cache/issues/135">#135</a>)</li> <li><a href="`def0926359`"><code>def0926</code></a> feat: add option to cache all crates (<a href="https://redirect.github.com/Swatinem/rust-cache/issues/137">#137</a>)</li> <li><a href="`827c240e23`"><code>827c240</code></a> fix: cache key dependency on installed packages (<a href="https://redirect.github.com/Swatinem/rust-cache/issues/138">#138</a>)</li> <li><a href="`5e9fae966f`"><code>5e9fae9</code></a> fix: cache restore failures (<a href="https://redirect.github.com/Swatinem/rust-cache/issues/136">#136</a>)</li> <li>Additional commits viewable in <a href="https://github.com/Swatinem/rust-cache/compare/v2.2.1...v2.4.0">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=Swatinem/rust-cache&package-manager=github_actions&previous-version=2.2.1&new-version=2.4.0)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) You can trigger a rebase of this PR by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) </details> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2023-06-05 07:58:52 +00:00
dependabot[bot]	a13ed4d0b0	Bump actions/setup-go from 3 to 4 Bumps [actions/setup-go](https://github.com/actions/setup-go) from 3 to 4. - [Release notes](https://github.com/actions/setup-go/releases) - [Commits](https://github.com/actions/setup-go/compare/v3...v4) --- updated-dependencies: - dependency-name: actions/setup-go dependency-type: direct:production update-type: version-update:semver-major ... Signed-off-by: dependabot[bot] <support@github.com>	2023-06-01 17:57:48 +00:00
dependabot[bot]	4cc2988482	Bump svenstaro/upload-release-action from 2.5.0 to 2.6.1 Bumps [svenstaro/upload-release-action](https://github.com/svenstaro/upload-release-action) from 2.5.0 to 2.6.1. - [Release notes](https://github.com/svenstaro/upload-release-action/releases) - [Changelog](https://github.com/svenstaro/upload-release-action/blob/master/CHANGELOG.md) - [Commits](https://github.com/svenstaro/upload-release-action/compare/2.5.0...2.6.1) --- updated-dependencies: - dependency-name: svenstaro/upload-release-action dependency-type: direct:production update-type: version-update:semver-minor ... Signed-off-by: dependabot[bot] <support@github.com>	2023-06-01 17:57:43 +00:00
dependabot[bot]	26c7e31f25	Bump Swatinem/rust-cache from 2.2.1 to 2.4.0 Bumps [Swatinem/rust-cache](https://github.com/Swatinem/rust-cache) from 2.2.1 to 2.4.0. - [Release notes](https://github.com/Swatinem/rust-cache/releases) - [Changelog](https://github.com/Swatinem/rust-cache/blob/master/CHANGELOG.md) - [Commits](https://github.com/Swatinem/rust-cache/compare/v2.2.1...v2.4.0) --- updated-dependencies: - dependency-name: Swatinem/rust-cache dependency-type: direct:production update-type: version-update:semver-minor ... Signed-off-by: dependabot[bot] <support@github.com>	2023-06-01 17:57:40 +00:00
meili-bors[bot]	b2dee07b5e	Merge #3783 3783: Improve SDK CI to choose the Docker image r=curquiza a=curquiza The point is to have the following "form" when running the SDK CI manually `nightly` is the default value if running the CI manually. <img width="1105" alt="Capture d’écran 2023-05-25 à 12 17 35" src="https://github.com/meilisearch/meilisearch/assets/20380692/87ae7123-efe8-4e7b-a99b-4a40aafa3f79"> Co-authored-by: curquiza <clementine@meilisearch.com>	2023-05-31 12:10:07 +00:00
meili-bors[bot]	d963b5f85a	Merge #3792 3792: fix the type of the document deletion by filter tasks r=dureuill a=irevoire # Pull Request ## Related issue Fixes https://github.com/meilisearch/meilisearch/issues/3791 ## What does this PR do? - Hide the deleteDocumentByFilter internal type from the users. Co-authored-by: Tamo <tamo@meilisearch.com>	2023-05-30 18:20:28 +00:00
Tamo	2acc3ec5ee	fix the type of the document deletion by filter tasks	2023-05-30 15:18:52 +02:00
Kerollmops	da04edff8c	Better use deserialize_unchecked_from to reduce the deserialization time	2023-05-30 14:58:30 +02:00
Tamo	85a80f4f4c	move the grafana dashboard to the assets directory and upload a basic prometheus scraper to help new users	2023-05-29 18:39:34 +02:00
Tamo	1213ec7164	update the dashboard once again	2023-05-29 18:37:55 +02:00
Tamo	f03d99690d	run the indexing fuzzer on every merge for as long as possible	2023-05-29 14:56:15 +02:00
meili-bors[bot]	0a7817a002	Merge #3786 3786: Consistently use wrapping add to avoid overflow in debug when query s… r=dureuill a=dureuill # Pull Request ## Related issue Fixes https://github.com/meilisearch/meilisearch/issues/3785 ## What does this PR do? - Some of the code paths would erroneously use the default addition operator that has the semantics that "overflow is an error, checked at runtime in debug" instead of the intended "overflow is expected" semantics that this code use (this code is using `u16::MAX` as a sentinel). This PR makes it so the wrapping add operator is used everywhere. Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-05-29 12:39:54 +00:00
Tamo	23a5b45ebf	drop the old fuzz file	2023-05-29 14:02:37 +02:00
Tamo	46fa99f486	make the fuzzer stops if an error occurs	2023-05-29 13:44:32 +02:00
Tamo	67a583bedf	handle the panic happening in milli	2023-05-29 13:39:26 +02:00
Tamo	99e9057684	rename the indexing fuzzer to fuzz-indexing so it doesn't collide with other binary name when being called from the root of the workspace	2023-05-29 13:07:06 +02:00
Tamo	8d40d300a5	rename the fuzzer to indexing	2023-05-29 12:37:24 +02:00
Tamo	6c6387d05e	move the fuzzer to its own crate	2023-05-29 12:27:39 +02:00
Louis Dureuil	1dfc4038ab	Add test that fails before PR and passes now	2023-05-29 11:58:26 +02:00
Louis Dureuil	73198179f1	Consistently use wrapping add to avoid overflow in debug when query starts with a separator	2023-05-29 11:54:12 +02:00
Tamo	51dce9e9d1	improve the dashboard slightly	2023-05-25 18:33:01 +02:00
Tamo	c9b65677bf	return the on disk size actually used by meilisearch	2023-05-25 18:30:30 +02:00
Tamo	35d5556f1f	prefix all the metrics by meilisearch_	2023-05-25 17:41:53 +02:00
Tamo	c433bdd1cd	add a view for the task queue in the metrics	2023-05-25 12:58:13 +02:00
curquiza	2db09725f8	Improve SDK CI to choose the Docker image	2023-05-25 12:22:35 +02:00
meili-bors[bot]	fdb23132d4	Merge #3781 3781: Revert "Improve docker cache" r=Kerollmops a=curquiza Reverts meilisearch/meilisearch#3566 because does not work as expected, and so I want to remove useless complexity from the CI and Dockerfile Co-authored-by: Clémentine U. - curqui <clementine@meilisearch.com>	2023-05-25 09:57:40 +00:00
Clémentine U. - curqui	11b95284cd	Revert "Improve docker cache"	2023-05-25 11:48:26 +02:00
Tamo	1b601f70c6	increase the bucketing of requests	2023-05-25 11:08:16 +02:00
meili-bors[bot]	8185731bbf	Merge #3779 3779: Add a cron test with disabled tokenization (with @roy9495) r=Kerollmops a=curquiza Replaces https://github.com/meilisearch/meilisearch/pull/3746 because of bors issue Co-authored-by: TATHAGATA ROY <98920199+roy9495@users.noreply.github.com> Co-authored-by: Clémentine U. - curqui <clementine@meilisearch.com>	2023-05-25 08:13:14 +00:00
Clémentine U. - curqui	840727d76f	Update .github/workflows/test-suite.yml	2023-05-25 10:07:59 +02:00
Clémentine U. - curqui	ead07d0b9d	Update .github/workflows/test-suite.yml	2023-05-25 10:07:52 +02:00
Clémentine U. - curqui	44f231d41e	Update .github/workflows/test-suite.yml	2023-05-25 10:07:45 +02:00
TATHAGATA ROY	3c5d1c93de	Added a cron test for disabled all-tokenization	2023-05-25 10:07:32 +02:00
meili-bors[bot]	087866d59f	Merge #3775 3775: Last error code changes on the new get/delete documents routes r=dureuill a=irevoire # Pull Request ## Related issue Fixes #3774 ## What does this PR do? Following the specification: https://github.com/meilisearch/specifications/pull/236 1. Get rid of the `invalid_document_delete_filter` and always use the `invalid_document_filter` 2. Introduce a new `missing_document_filter` instead of returning `invalid_document_delete_filter` (that’s consistent with all the other routes that have a mandatory parameter) 3. Always return the `original_filter` in the details (potentially set to `null`) instead of hiding it if it wasn’t used Co-authored-by: Tamo <tamo@meilisearch.com>	2023-05-24 10:07:41 +00:00
Tamo	9111f5176f	get rid of the invalid document delete filter in favor of the invalid document filter	2023-05-24 11:53:16 +02:00
Tamo	b9dd092a62	make the details return null in the originalFilter field if no filter was provided + add a big test on the details	2023-05-24 11:48:22 +02:00
Tamo	ca99bc3188	implement the missing document filter error code when deleting documents	2023-05-24 11:29:20 +02:00
Tamo	57d53de402	Increase the number of buckets	2023-05-24 10:47:15 +02:00
meili-bors[bot]	2e49d6aec1	Merge #3768 3768: Fix bugs in graph-based ranking rules + make `words` a graph-based ranking rule r=dureuill a=loiclec This PR contains three changes: ## 1. Don't call the `words` ranking rule if the term matching strategy is `All` This is because the purpose of `words` is only to remove nodes from the query graph. It would never do any useful work when the matching strategy was `All`. Remember that the universe was already computed before by computing all the docids corresponding to the "maximally reduced" query graph, which, in the case of `All`, is equal to the original graph. ## 2. The `words` ranking rule is replaced by a graph-based ranking rule. This is for three reasons: 1. performance: graph-based ranking rules benefit from a lot of optimisations by default, which ensures that they are never too slow. The previous implementation of `words` could call `compute_query_graph_docids` many times if some words had to be removed from the query, which would be quite expensive. I was especially worried about its performance in cases where it is placed right after the `sort` ranking rule. Furthermore, `compute_query_graph_docids` would clone a lot of bitmaps many times unnecessarily. 2. consistency: every other ranking rule (except `sort`) is graph-based. It makes sense to implement `words` like that as well. It will automatically benefit from all the features, optimisations, and bug fixes that all the other ranking rules get. 3. surfacing bugs: as the first ranking rule to be called (most of the time), I'd like `words` to behave the same as the other ranking rules so that we can quickly detect bugs in our graph algorithms. This actually already happened, which is why this PR also contains a bug fix. ## 3. Fix the `update_all_costs_before_nodes` function It is a bit difficult to explain what was wrong, but I'll try. The bug happened when we had graphs like: <img width="730" alt="Screenshot 2023-05-16 at 10 58 57" src="https://github.com/meilisearch/meilisearch/assets/6040237/40db1a68-d852-4e89-99d5-0d65757242a7"> and we gave the node `is` as argument. Then, we'd walk backwards from the node breadth-first. We'd update the costs of: 1. `sun` 2. `thesun` 3. `start` 4. `the` which is an incorrect order. The correct order is: 1. `sun` 2. `thesun` 3. `the` 4. `start` That is, we can only update the cost of a node when all of its successors have either already been visited or were not affected by the update to the node passed as argument. To solve this bug, I factored out the graph-traversal logic into a `traverse_breadth_first_backward` function. Co-authored-by: Loïc Lecrenier <loic.lecrenier@me.com> Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-05-23 13:28:08 +00:00
Louis Dureuil	51043f78f0	Remove trailing whitespace	2023-05-23 15:27:25 +02:00
Louis Dureuil	a490a11325	Add explanatory comment on the way we're recomputing costs	2023-05-23 15:24:24 +02:00
Tamo	002f42875f	fix the fuzzer	2023-05-23 11:42:40 +02:00
Tamo	22213dc604	push the fuzzer	2023-05-23 09:14:26 +02:00
Tamo	602ad98cb8	improve the way we handle the fsts	2023-05-22 11:15:14 +02:00
Tamo	7f619ff0e4	get rids of the now unused soft_deletion_used parameter	2023-05-22 10:33:49 +02:00
Tamo	4391cba6ca	fix the addition + deletion bug	2023-05-17 18:28:57 +02:00
Tamo	d7ddf4925e	Revert "Disable autobatching of additions and deletions" This reverts commit `a94e78ffb0`.	2023-05-17 14:25:50 +02:00
meili-bors[bot]	101f5a20d2	Merge #3757 3757: Adjust the cost of edges in the `position` ranking rule by bucketing positions more aggressively r=loiclec a=loiclec This PR significantly improves the performance of the `position` ranking rule when: 1. a query contains many words 2. the `position` ranking rule needs to be called many times 3. the score of the documents according to `position` is high These conditions greatly increase: 1. the number of edge traversals that are needed to find a valid path from the `start` node to the `end` node 2. the number of edges that need to be deleted from the graph, and therefore the number of times that we need to recompute all the possible costs from START to END As a result, a majority of the search time is spent in `visit_condition`, `visit_node`, and `update_all_costs_before_node`. This is frustrating because it often happens when the "universe" given to the rule consists of only a handful of document ids. By limiting the number of possible edges between two nodes from `20` to `10`, we: 1. reduce the number of possible costs from START to END 2. reduce the number of edges that will be deleted 3. make it faster to update the costs after deleting an edge 4. reduce the number of buckets that need to be computed In terms of relevancy, I don't think we lose or gain much. We still prefer terms that are in a lower positions, with decreasing precision as we go further. The previous choice of bucketing wasn't chosen in a principled way, and neither is this one. They both "feel" right to me. Co-authored-by: Loïc Lecrenier <loic.lecrenier@me.com> Co-authored-by: meili-bors[bot] <89034592+meili-bors[bot]@users.noreply.github.com>	2023-05-17 11:43:59 +00:00
meili-bors[bot]	6ce1ce77e6	Merge #3738 3738: Add analytics on the get documents resource r=dureuill a=irevoire # Pull Request ## Related issue Fixes https://github.com/meilisearch/meilisearch/issues/3737 Related spec https://github.com/meilisearch/specifications/pull/234 ## What does this PR do? Add the analytics for the following routes: - `GET` - `/indexes/:uid/documents` - `GET` - `/indexes/:uid/documents/:doc_id` - `POST` - `/indexes/:uid/documents/fetch` These analytics are aggregated between two events: - `Documents Fetched GET` - `Documents Fetched POST` That shares the same payload: Property name \| Description \| Example \| \|---------------\|-------------\|---------\| \| `requests.total_received` \| Total number of request received in this batch \| 325 \| \| `per_document_id` \| `false` \| false \| \| `per_filter` \| `true` if `POST /indexes/:indexUid/documents/fetch` endpoint was used with a filter in this batch, otherwise `false` \| false \| \| `pagination.max_limit` \| Highest value given for the `limit` parameter in this batch \| 60 \| \| `pagination.max_offset` \| Highest value given for the `offset` parameter in this batch \| 1000 \| Co-authored-by: Tamo <tamo@meilisearch.com>	2023-05-16 19:37:41 +00:00
Loïc Lecrenier	ec8f685d84	Fix bug in cheapest path algorithm	2023-05-16 17:01:30 +02:00
Loïc Lecrenier	5758268866	Don't compute split_words for phrases	2023-05-16 17:01:18 +02:00
meili-bors[bot]	4d037e6693	Merge #3759 3759: Invalid error code when parsing filters r=dureuill a=irevoire # Pull Request ## Related issue Fixes https://github.com/meilisearch/meilisearch/issues/3753 ## What does this PR do? Fix the error code in case the error comes from the evaluate of the filter for the get, fetch and delete documents routes. Co-authored-by: Tamo <tamo@meilisearch.com>	2023-05-16 12:55:06 +00:00
Tamo	96da5130a4	fix the error code in case of not filterable attributes on the get / delete documents by filter routes	2023-05-16 13:56:18 +02:00
Loïc Lecrenier	3e19702de6	Update snapshot tests	2023-05-16 12:22:46 +02:00
meili-bors[bot]	1e762d151f	Merge #3755 3755: Re-add final dot r=curquiza a=ManyTheFish I removed the final dot of the error message in my last PR, this one re-adds it. related to https://github.com/meilisearch/meilisearch/pull/3749 > Oups 😬 Co-authored-by: ManyTheFish <many@meilisearch.com>	2023-05-16 10:10:58 +00:00
Tamo	0b38f211ac	test the new introduced route	2023-05-16 12:07:44 +02:00
Loïc Lecrenier	f6524a6858	Adjust costs of edges in position ranking rule To ensure good performance	2023-05-16 11:28:56 +02:00
meili-bors[bot]	65ad8cce36	Merge #3741 3741: Add ngram support to the highlighter r=ManyTheFish a=loiclec This PR fixes a bug introduced by the search refactor, where ngrams were not highlighted. The solution was to add the ngrams to the vector of `LocatedQueryTerm` that is given to the `MatchingWords` structure. Co-authored-by: Loïc Lecrenier <loic.lecrenier@me.com>	2023-05-16 09:03:31 +00:00
ManyTheFish	42650f82e8	Re-add final dot	2023-05-16 10:57:26 +02:00
Loïc Lecrenier	a37da36766	Implement `words` as a graph-based ranking rule and fix some bugs	2023-05-16 10:42:11 +02:00
Loïc Lecrenier	85d96d35a8	Highlight ngram matches as well	2023-05-16 10:39:36 +02:00
Filip Bachul	64b11f45d7	fix test name	2023-05-16 09:24:49 +02:00
meili-bors[bot]	bf66e97b48	Merge #3749 3749: Fix back: sort error message r=ManyTheFish a=ManyTheFish This PR reintroduces the error message modified in https://github.com/meilisearch/milli/pull/375. However, this added double-quotes around `sort` in the message. I don't think another message contains double-quotes, so I have added a separate commit replacing the double-quotes with back-ticks, which seems more consistent with the other error messages, this last change can be reverted easily. ## Detailed changes #### v1.2-rc0 ``` The sort ranking rule must be specified in the ranking rules settings to use the sort parameter at search time. ``` #### [Reintroduce fix (previous and expected behavior)](`23d1c86825`) ``` You must specify where "sort" is listed in the rankingRules setting to use the sort parameter at search time ``` #### [Replace double-quotes with back-ticks (my suggestion)](`4d691d071a`) ``` You must specify where `sort` is listed in the rankingRules setting to use the sort parameter at search time ``` ## Related Fixes #3722 ## Reviewers - technical review: `@irevoire` - to validate the replacement: `@macraig` Co-authored-by: ManyTheFish <many@meilisearch.com>	2023-05-15 14:55:51 +00:00
meili-bors[bot]	a7ea5ec748	Merge #3651 3651: Use the writemap flag to reduce the memory usage r=irevoire a=Kerollmops This draft PR is showing some stats about the memory usage of Meilisearch when [the LMDB `MDB_WRITEMAP` flag](`3947014aed/libraries/liblmdb/lmdb.h (L573-L581)`) is enabled and when it is not. As you can see there is a reduction of about 50% of the memory usage pick. The dataset used was [the Wikipedia one](https://www.notion.so/meilisearch/Wikipedia-8b1486e4b17547c5bda485d2d97767a0) with the first 30 000 first CSV documents without settings. This PR depends on https://github.com/meilisearch/heed/pull/168. I just [opened a discussion](https://github.com/meilisearch/product/discussions/652) for people to understand the tradeoffs and give their feedback. - [x] Create an experiment flag `--experimental-reduce-indexing-memory-usage`. - [x] Add it to the config file. - [x] Explain the tradeoff and copy/link the LMDB documentation in the help message. - [x] Add analytics about the experimental flag. - [x] Document that this flag cannot be used on Windows, ~~or hide it~~. <details> <summary>The command I used to run the tests</summary> #### Sign the binary to be able to use Instruments / xcrun ```sh codesign -s - -f --entitlements ~/ent.plist target/release/meilisearch ``` where `ent.plist` contains: ```xml <?xml version="1.0" encoding="UTF-8"?> <!DOCTYPE plist PUBLIC "-//Apple//DTD PLIST 1.0//EN" "http://www.apple.com/DTDs/PropertyList-1.0.dtd"> <plist version="1.0"> <dict> <key>com.apple.security.get-task-allow</key> <true/> </dict> </plist> ``` #### Run Meilisearch in measure-mode ```sh xcrun xctrace record --template 'Allocations' --launch -- target/release/meilisearch --max-indexing-memory 0MiB ``` #### Send the wiki dataset available on notion.so / Public ```sh for f in 0.csv 15000.csv; do echo sending $f; xh 'localhost:7700/indexes/wiki/documents' 'content-type:text/csv' `@$f;` done ``` #### Wait for the task to finish ```sh watch --color xh --pretty all 'localhost:7700/tasks?statuses=processing' ``` </details> Keep in mind that I tested that with the Instruments Apple tools on an iMac 5k 2019. More benchmarks must be done, especially on the indexation speed, as the flag is told to slow down writing into databases bigger that the amount of memory. On the left Meilisearch is running without the flag. On the right, it is running with the flag. <p align="center"> <img align="left" width="45%" alt="Instrument showing the memory usage of Meilisearch without the MDB_WRITEMAP flag" src="https://user-images.githubusercontent.com/3610253/234299524-7607f1df-6fc1-45d3-bd3d-4f9388002857.png"> <img align="right" width="45%" alt="Instrument showing the memory usage of Meilisearch with the MDB_WRITEMAP flag" src="https://user-images.githubusercontent.com/3610253/234299534-6cc3ae58-8bd9-426c-aa79-4c78f9e88b94.png"> </p> Co-authored-by: Kerollmops <clement@meilisearch.com> Co-authored-by: Clément Renault <clement@meilisearch.com>	2023-05-15 14:10:07 +00:00
Kerollmops	dc7ba77e57	Add the option in the config file	2023-05-15 16:07:43 +02:00
Clément Renault	13f870e993	Fix typos and documentation issues	2023-05-15 15:11:45 +02:00
Kerollmops	1a79fd0c3c	Use the new heed v0.12.6	2023-05-15 11:42:30 +02:00
Kerollmops	f759ec7fad	Expose a flag to enable the MDB_WRITEMAP flag	2023-05-15 11:38:43 +02:00
ManyTheFish	4d691d071a	Change double-quotes by back-ticks in sort error message	2023-05-15 11:10:36 +02:00
ManyTheFish	23d1c86825	Re-introduce the sort error message fix	2023-05-15 11:07:23 +02:00
Kerollmops	c4a40e7110	Use the writemap flag to reduce the memory usage	2023-05-15 10:15:33 +02:00
Filip Bachul	e68d86d6b6	tests: add unit test for `PayloadTooLarge`	2023-05-11 20:51:10 +02:00
meili-bors[bot]	e01980c6f4	Merge #3739 3739: fix: update `payload_too_large` error message to include human readable maximum acceptable payload size r=Kerollmops a=cymruu # Pull Request ## Related issue Fixes #3736 ## What does this PR do? - update `payload_too_large` error message as requested in ticket ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Filip Bachul <filipbachul@gmail.com>	2023-05-11 09:37:19 +00:00
Filip Bachul	25209a3590	introduce `remaining` field in `Payload`	2023-05-10 20:55:18 +02:00
Filip Bachul	3064ea6495	fix: update payload_too_large error message to include human readable maximum acceptable payload size	2023-05-10 18:16:59 +02:00
Tamo	46ec8a97e9	rename the analytics according to the spec	2023-05-10 14:28:30 +02:00
Tamo	c42a65a297	Update meilisearch/src/analytics/segment_analytics.rs Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-05-10 14:28:30 +02:00
Tamo	d08f8690d2	add analytics on the get documents resource	2023-05-10 14:28:30 +02:00
meili-bors[bot]	ad5f25d880	Merge #3742 3742: Compute split words derivations of terms that don't accept typos r=ManyTheFish a=loiclec Allows looking for the split-word derivation for short words in the user's query (like `the -> "t he"` or `door -> do or`) as well as for 3grams. Co-authored-by: Loïc Lecrenier <loic.lecrenier@me.com>	2023-05-10 12:12:52 +00:00
Loïc Lecrenier	4d352a21ac	Compute split words derivations of terms that don't accept typos	2023-05-10 13:31:19 +02:00
meili-bors[bot]	918ce1dd67	Merge #3731 3731: Move comments above keys in config.toml r=curquiza a=jirutka The current style is very unusual, confusing and breaks compatibility with tools for parsing config files including comments. Everyone writes comments above the items to which they refer (maybe except pythonists), so let's stick to that. Co-authored-by: Jakub Jirutka <jakub@jirutka.cz>	2023-05-09 09:24:36 +00:00
meili-bors[bot]	4a4210c116	Merge #3734 3734: Update version for the next release (v1.2.0) in Cargo.toml r=curquiza a=meili-bot ⚠️ This PR is automatically generated. Check the new version is the expected one and Cargo.lock has been updated before merging. Co-authored-by: curquiza <curquiza@users.noreply.github.com>	2023-05-09 07:35:48 +00:00
curquiza	3533d4f2bb	Update version for the next release (v1.2.0) in Cargo.toml	2023-05-08 17:52:33 +00:00
Loïc Lecrenier	3625389057	Highlight ngram matches as well	2023-05-08 15:35:41 +02:00
meili-bors[bot]	eace6df91b	Merge #3726 3726: Fix prefix highlighting r=loiclec a=ManyTheFish The prefix queries were not properly highlighted, this PR now highlights only the start of a word when it matched with a prefix Co-authored-by: ManyTheFish <many@meilisearch.com> Co-authored-by: Loïc Lecrenier <loic.lecrenier@me.com>	2023-05-08 07:46:46 +00:00
Loïc Lecrenier	83ab8cf4e5	Remove dbg!(..) expression in highlighter tests	2023-05-08 09:45:23 +02:00
Jakub Jirutka	8095f21999	Move comments above keys in config.toml The current style is very unusual, confusing and breaks compatibility with tools for parsing config files including comments. Everyone writes comments above the items to which they refer (maybe except pythonists), so let's stick to that.	2023-05-06 18:10:54 +02:00
ManyTheFish	cd2573fcc3	Fix prefix highlighting	2023-05-04 16:53:50 +02:00
meili-bors[bot]	9f7981df28	Merge #3687 3687: Allow to disable specialized tokenizations (again) r=Kerollmops a=jirutka In PR #2773, I added the `chinese`, `hebrew`, `japanese` and `thai` feature flags to allow melisearch to be built without huge specialed tokenizations that took up 90% of the melisearch binary size. Unfortunately, due to some recent changes, this doesn't work anymore. The problem lies in excessive use of the `default` feature flag, which infects the dependency graph. Instead of adding `default-features = false` here and there, it's easier and more future-proof to not declare `default` in `milli` and `meilisearch-types`. I've renamed it to `all-tokenizers`, which also makes it a bit clearer what it's about. Co-authored-by: Jakub Jirutka <jakub@jirutka.cz>	2023-05-04 14:48:01 +00:00
Jakub Jirutka	e615fa5ec6	Fix unused_imports warning in milli when japanese is not enabled	2023-05-04 15:46:11 +02:00
Jakub Jirutka	13f1277637	Allow to disable specialized tokenizations (again) In PR #2773, I added the `chinese`, `hebrew`, `japanese` and `thai` feature flags to allow melisearch to be built without huge specialed tokenizations that took up 90% of the melisearch binary size. Unfortunately, due to some recent changes, this doesn't work anymore. The problem lies in excessive use of the `default` feature flag, which infects the dependency graph. Instead of adding `default-features = false` here and there, it's easier and more future-proof to not declare `default` in `milli` and `meilisearch-types`. I've renamed it to `all-tokenizers`, which also makes it a bit clearer what it's about.	2023-05-04 15:45:40 +02:00
meili-bors[bot]	4919774f2e	Merge #3570 3570: Get documents by filter r=irevoire a=dureuill # Pull Request ## Related issue Associated spec: https://github.com/meilisearch/specifications/pull/234 None really, this is more of an extension of #3477: since after this issue we'll be able to delete documents by filter, it makes sense to also be able to get documents by filter. ## What does this PR do? ### User standpoint - Add a new `filter` URL parameter to `GET /indexes/{:indexUid}/documents` and a new `POST /indexes/{:indexUid}/documents/fetch` route with the same `offset, limit, fields, filter` ### Implementation standpoint - Add a new `Index::iter_documents` method to iterate on a set of documents rather than return a vector of these documents. - Rewrite the other `Index::*documents` methods to use the new `Index::iter_documents` method. ## Usage <details> <summary> Sample request and response </summary> ``` curl -X POST 'http://localhost:7700/indexes/index-1101/documents/fetch' -H 'Content-Type: application/json' --data-binary '{ "filter": "genres = Comedy", "limit": 3, "offset": 8000}' \| jsonxf ``` ```json { "results": [ { "id": 326126, "title": "Bad Exorcists", "overview": "A trio of awkward teens intend to win a horror festival by making their own movie, but wind up getting their actress possessed in the process.", "genres": [ "Horror", "Comedy" ], "poster": "https://image.tmdb.org/t/p/w500/lwd65kPbjFacAw3QSXiwSsW6cFU.jpg", "release_date": 1425081600 }, { "id": 326215, "title": "Ooops! Noah is Gone...", "overview": "It's the end of the world. A flood is coming. Luckily for Dave and his son Finny, a couple of clumsy Nestrians, an Ark has been built to save all animals. But as it turns out, Nestrians aren't allowed. Sneaking on board with the involuntary help of Hazel and her daughter Leah, two Grymps, they think they're safe. Until the curious kids fall off the Ark. Now Finny and Leah struggle to survive the flood and hungry predators and attempt to reach the top of a mountain, while Dave and Hazel must put aside their differences, turn the Ark around and save their kids. It's definitely not going to be smooth sailing.", "genres": [ "Animation", "Adventure", "Comedy", "Family" ], "poster": "https://image.tmdb.org/t/p/w500/gEJXHgpiKh89Vwjc4XUY5CIgUdB.jpg", "release_date": 1427328000 }, { "id": 326241, "title": "For Here or to Go?", "overview": "An aspiring Indian tech entrepreneur in the Silicon Valley finds himself unexpectedly battling the bizarre American immigration system to keep his dream alive or prepare to return home forever.", "genres": [ "Drama", "Comedy" ], "poster": "https://image.tmdb.org/t/p/w500/ff8WaA7ItBgl36kdT232i0d0Fnq.jpg", "release_date": 1490918400 } ], "offset": 8000, "limit": 3, "total": 9331 } ``` <img width="1348" alt="Capture d’écran 2023-03-08 à 10 09 04" src="https://user-images.githubusercontent.com/41078892/223670905-6932b79b-f9b8-4a41-b59e-be2171705b7d.png"> </details> # Draft status - [ ] Route naming: having one route be `GET /indexes/{:indexUid}/documents` and the other `POST /indexes/{:indexUid}/documents/fetch` is suboptimal (also, technically a breaking change for documents with `fetch` as uid?), but `POST /indexes/{:indexUid}/documents` is already used to insert documents. Co-authored-by: Louis Dureuil <louis@meilisearch.com> Co-authored-by: Tamo <tamo@meilisearch.com>	2023-05-04 12:54:26 +00:00
Tamo	a3da680ce6	Update meilisearch/tests/documents/errors.rs Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-05-04 14:51:17 +02:00
Tamo	11e394dba1	merge the document fetch and get error codes	2023-05-04 15:39:49 +02:00
Tamo	469d2f2a9c	fix the fields field of the POST fetch document API	2023-05-04 15:34:09 +02:00
Tamo	ce6507d20c	improve the test of the get document by filter	2023-05-04 15:34:09 +02:00
Tamo	b92da5d15a	add a big test on the get document by filter of the get route	2023-05-04 15:34:09 +02:00
Tamo	ed3dfbe729	add error codes and tests	2023-05-04 15:34:08 +02:00
Louis Dureuil	441641397b	Implement document get with filters	2023-05-04 15:32:34 +02:00
Louis Dureuil	a35d3fc708	Add Index::iter_documents	2023-05-04 15:31:54 +02:00
Louis Dureuil	745c1a2668	Make parse_filter pub	2023-05-04 15:31:53 +02:00
meili-bors[bot]	a95128df6b	Merge #3550 3550: Delete documents by filter r=irevoire a=dureuill # Prototype `prototype-delete-by-filter-0` Usage: A new route is available under `POST /indexes/{index_uid}/documents/delete` that allows you to delete your documents by filter. The expected payload looks like that: ```json { "filter": "doggo = bernese", } ``` It'll then enqueue a task in your task queue that'll delete all the documents matching this filter once it's processed. Here is an example of the associated details; ```json "details": { "deletedDocuments": 53, "originalFilter": "\"doggo = bernese\"" } ``` ---------- # Pull Request ## Related issue Related to https://github.com/meilisearch/meilisearch/issues/3477 ## What does this PR do? ### User standpoint - Modifies the `/indexes/{:indexUid}/documents/delete-batch` route to accept either the existing array of documents ids, or a JSON object with a `filter` field representing a filter to apply. If that latter variant is used, any document matching the filter will be deleted. ### Implementation standpoint - (processing time version) Adds a new BatchKind that is not autobatchable and that performs the delete by filter - Reuse the `documentDeletion` task with a new `originalFilter` detail that replaces the `providedIds` detail. ## Example <details> <summary>Sample request, response and task result</summary> Request: ``` curl \ -X POST 'http://localhost:7700/indexes/index-10/documents/delete-batch' \ -H 'Content-Type: application/json' \ --data-binary '{ "filter" : "mass = 600"}' ``` Response: ``` { "taskUid": 3902, "indexUid": "index-10", "status": "enqueued", "type": "documentDeletion", "enqueuedAt": "2023-02-28T20:50:31.667502Z" } ``` Task log: ```json { "uid": 3906, "indexUid": "index-12", "status": "succeeded", "type": "documentDeletion", "canceledBy": null, "details": { "deletedDocuments": 3, "originalFilter": "\"mass = 600\"" }, "error": null, "duration": "PT0.001819S", "enqueuedAt": "2023-03-07T08:57:20.11387Z", "startedAt": "2023-03-07T08:57:20.115895Z", "finishedAt": "2023-03-07T08:57:20.117714Z" } ``` </details> ## Draft status - [ ] Error handling - [ ] Analytics - [ ] Do we want to reuse the `delete-batch` route in this way, or create a new route instead? - [ ] Should the filter be applied at request time or when the deletion task is processed? - The first commit in this PR applies the filter at request time, meaning that even if a document is modified in a way that no longer matches the filter in a later update, it will be deleted as long as the deletion task is processed after that update. - The other commits in this PR apply the filter only when the asynchronous deletion task is processed, meaning that documents that match the filter at processing time are deleted even if they didn't match the filter at request time. - [ ] If keeping the filter at request time, find a more elegant way to recover the user document ids from the internal document ids. The current way implemented in the first commit of this PR involves getting all the documents matching the filter, looking for the value of their primary key, and turning it into a string by copy-pasting routines found in milli... - [ ] Security consideration, if any - [ ] Fix the tests (but waiting until product questions are resolved) - [ ] Add delete by filter specific tests Co-authored-by: Louis Dureuil <louis@meilisearch.com> Co-authored-by: Tamo <tamo@meilisearch.com>	2023-05-04 10:44:41 +00:00
meili-bors[bot]	e0537c3870	Merge #3720 3720: Change links of docs everywhere r=curquiza a=curquiza Completely fixes #3668 Co-authored-by: curquiza <clementine@meilisearch.com>	2023-05-04 10:07:41 +00:00
meili-bors[bot]	da220294f6	Merge #3639 3639: Add a dedicated error variant for planned failures in index scheduler tests r=Kerollmops a=Sufflope # Pull Request ## Related issue Fixes #3086 ## What does this PR do? - Add a dedicated test variant in test cfg to avoid reusing a misleading existing error ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Jean-Sébastien Bour <jean-sebastien@bour.name>	2023-05-04 09:33:57 +00:00
meili-bors[bot]	78e611f282	Merge #3693 3693: Implement the auto deletion of tasks r=dureuill a=irevoire Fixes https://github.com/meilisearch/meilisearch/issues/3622 This PR should be the definite fix for #3622. It adds a limit (1M) to the maximum number of tasks the task queue can hold. Once the task queue reaches this limit (1M of tasks are in the task queue, whatever their status is), meilisearch will schedule a task deletion that tries to delete the oldest 100k tasks. If meilisearch can't delete 100k tasks because some of them are not yet finished, it will delete as many tasks as possible. Once the limit is reached, you're still able to register new tasks. The engine will only stop you from adding new tasks once [the other hard limit](https://github.com/meilisearch/meilisearch/pull/3659) of 10GiB of tasks is reached (that's between 5M and 15M of tasks depending on your workflow). ------- Technically; - We only try to schedule our task deletion when calling the tick function but before creating a new batch. This means we never enqueue a task we're not going to process ~right away. - If our task deletion doesn't delete anything, we don't enqueue it and log a warn the user that the engine is not working properly Co-authored-by: Tamo <tamo@meilisearch.com> Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-05-04 08:30:22 +00:00
Louis Dureuil	d8381eb790	Fix originalFilter	2023-05-04 10:07:59 +02:00
Louis Dureuil	b212aef5db	add one nanosecond to generated filter so as to generate a filter that would have matched the last task to delete	2023-05-04 09:56:48 +02:00
meili-bors[bot]	6bf66f35be	Merge #3721 3721: Use new bors URL of our self hosted bors instance r=curquiza a=curquiza Co-authored-by: curquiza <clementine@meilisearch.com>	2023-05-04 07:53:39 +00:00
Louis Dureuil	52ab114f6c	Fix test on macOS: 50 tasks would result in the test consistently failing on a local macOS	2023-05-04 00:06:49 +02:00
Tamo	dcbfecf42c	make the generated filter valid	2023-05-04 00:06:49 +02:00
Tamo	9ca6f59546	Update index-scheduler/src/lib.rs Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-05-04 00:06:49 +02:00
Tamo	aa7537a11e	make the autodeletion work with a fixed number of tasks and update the tests	2023-05-04 00:06:49 +02:00
Tamo	972bb2831c	log when meilisearch need to delete tasks	2023-05-04 00:06:49 +02:00
Tamo	f9ddd32545	implement the auto-deletion of tasks	2023-05-04 00:06:49 +02:00
Louis Dureuil	d5059520aa	Fix typo	2023-05-03 22:27:03 +02:00
Louis Dureuil	1c3642c9b2	Fix deletion per filter analytics	2023-05-03 22:26:51 +02:00
Tamo	d2d2bacaf2	add a test on the complex filter	2023-05-03 20:07:08 +02:00
curquiza	30edba3497	Update links of the docs	2023-05-03 19:14:57 +02:00
Louis Dureuil	84e7bd9342	Fix test after rebase on filter additions	2023-05-03 17:51:28 +02:00
Louis Dureuil	2b74e4d116	Fix test	2023-05-03 17:41:50 +02:00
Tamo	b5fe0b2b07	fix the details	2023-05-03 17:41:50 +02:00
Tamo	0f0cd2d929	handle the array of array form of filter in the dumps	2023-05-03 17:41:50 +02:00
Tamo	fc8c1d118d	fix the analytics	2023-05-03 17:41:50 +02:00
Tamo	0548ab9038	create and use the error code	2023-05-03 17:41:50 +02:00
Tamo	143acb9cdc	update the tests	2023-05-03 17:41:49 +02:00
Tamo	4b92f1b269	wip	2023-05-03 17:41:49 +02:00
Tamo	c12a1cd956	test all the error messages	2023-05-03 17:41:49 +02:00
Tamo	8af8aa5a33	add a test	2023-05-03 17:41:49 +02:00
Tamo	6df2ba93a9	remove one useless txn	2023-05-03 17:41:49 +02:00
Louis Dureuil	3680a6bf1e	extract impl to a function	2023-05-03 17:41:49 +02:00
Louis Dureuil	732c52093d	Processing time without autobatching implementation	2023-05-03 17:41:48 +02:00
Louis Dureuil	05cc463fbc	Draft implementation of filter support for /delete-by-batch route	2023-05-03 17:41:48 +02:00
meili-bors[bot]	1afde4fea5	Merge #3542 3542: Refactor of the search algorithms r=dureuill a=loiclec This PR refactors a large part of the search logic (related to https://github.com/meilisearch/meilisearch/issues/3547) - The "query tree" is replaced by a "query graph", which describes the different ways in which the search query can be interpreted and precomputes the word derivations for each query term. Example: <img width="1162" alt="Screenshot 2023-02-27 at 10 26 50" src="https://user-images.githubusercontent.com/6040237/221525270-87917cc0-60d1-473f-847f-2c5a7de9e370.png"> - The control flow between the ~criterions~ ranking rules is managed in a single place instead of being independently implemented by each ranking rule. - The set of document candidates is determined greedily from the beginning. It is often referred as the "universe" in the code. - The ranking rules `proximity`, `attribute`, `typo`, and (maybe) `exactness` are or will be implemented using a K-shortest path graph algorithm. This minimises the number of database and bitmap operations we need to do to compute each ranking rule bucket. It also simplifies the code a lot since a lot of ranking rules will share a large part of their implementation. - Pointers to database values are stored in a cache to avoid searching in the LMDB databases needlessly. - The result of some roaring bitmap operations are also stored in a cache, although we'll need to measure the memory pressure this puts on the system and maybe deactivate this cache later on. - Search requests can be visually logged and debugged in tests. TODO: - [ ] Reintroduce search benchmarks - [x] Implement `disableOnWords` and `disableOnAttributes` settings of typo tolerance - [x] Implement "exhaustive number of hits - [x] Implement `attribute` ranking rule - [x] Indexing changes: split into `word_fid_docids` and `word_position_docids` (with bucketed position) - [x] Ranking rule implementations - [ ] Implement `exactness` ranking rule - [x] Initial implementation - [ ] Correct implementation when followed by `Words` - [ ] Implement `geosort` ranking rule - [ ] Add tests - [x] Typo tolerance `disableOnWords`/`disableOnAttributes` - [ ] Geosort - [x] Exactness - [ ] Attribute/Position - [ ] Interactions between ranking rules: - [x] Typo/Proximity/Attribute not preceded by Words - [x] Exactness not preceded by Words - [x] Exactness -> Words (+ check universe correctness) - [x] Exactness -> Typo, etc. - [ ] Sort -> Words (performance tests) - [ ] Attribute/Position -> Typo - [ ] Attribute/Position -> Proximity - [x] Typo -> Exactness - [x] Typo -> Proximity - [x] Proximity -> Typo - [x] Words - [x] Typo - [x] Proximity - [x] Sort - [x] Ngrams - [x] Split words - [x] Ngram + Split Words - [x] Term matching strategy - [x] Distinct attribute - [x] Phrase Search - [x] Placeholder search - [x] Highlighter - [x] Limit the number of word derivations in a search query - [x] Compute the initial universe correctly according to the terms matching strategy - [x] Implement placeholder search - [x] Get the list of ranking rules from the settings - [x] Implement `distinct` - [x] Determine what to do when one of `attribute`, `proximity`, `typo`, or `exactness` is placed before `words` - [x] Make sure the correct number of allowed typos is used for each word, including the prefix one - [x] Make sure stop words are treated correctly (e.g. correct position in query graph), including in phrases - [x] Support phrases correctly - [x] Support synonyms - [x] Support split words - [x] Support combination of ngram + split-words (e.g. `whiteh orse` -> `"white horse"`) - [x] Implement `typo` ranking rule - [x] Implement `sort` ranking rule - [x] Use existing `Search` interface to use the new search algorithms - [x] Remove old code Co-authored-by: Loïc Lecrenier <loic.lecrenier@me.com>	2023-05-03 13:42:51 +00:00
Louis Dureuil	f8f190cd40	Update exactness tests following charabia camelCase tokenization	2023-05-03 14:45:09 +02:00
Louis Dureuil	3a408e8287	Increase map size for tests following charabia camelCase tokenization	2023-05-03 14:44:48 +02:00
Louis Dureuil	d3e5b10e23	fix nb of dbs	2023-05-03 14:11:20 +02:00
Louis Dureuil	1aaf24ccbf	Cargo fmt	2023-05-03 12:21:58 +02:00
Louis Dureuil	90bc230820	Merge remote-tracking branch 'origin/main' into search-refactor Conflicts \| resolution ----------\|----------- Cargo.lock \| added mimalloc Cargo.toml \| took origin/main version milli/src/search/criteria/exactness.rs \| deleted after checking it was only clippy changes milli/src/search/query_tree.rs \| deleted after checking it was only clippy changes	2023-05-03 12:19:06 +02:00
Louis Dureuil	342c4ff85d	geosort: Remove rtree unwrap	2023-05-03 09:52:16 +02:00
Tamo	c85392ce40	make the descendent geosort fast	2023-05-03 09:13:12 +02:00
Tamo	8875d24a48	deserialize the rtree only when its needed, and keep it in memory once it has been deserialized	2023-05-03 09:13:12 +02:00
Tamo	c470b67fa2	revamp the test to use execute_iterative_and_rtree_returns_the_same	2023-05-03 09:13:12 +02:00
meili-bors[bot]	c0e081cd98	Merge #3702 #3710 3702: Update charabia v0.7.2 r=curquiza a=ManyTheFish fixes #3701 fixes #3689 fixes #3285 3710: Updated messages pointing to the docs website r=curquiza a=roy9495 # Pull Request Fixes partially #3668 ## What does this PR do? - ...Any messages referencing this docs site https://docs.meilisearch.com has been changed to this docs site https://meilisearch.com/docs . Thanks. ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: ManyTheFish <many@meilisearch.com> Co-authored-by: TATHAGATA ROY <98920199+roy9495@users.noreply.github.com>	2023-05-02 17:27:57 +00:00
Louis Dureuil	b60840ebff	Remove self.iterating from words	2023-05-02 18:54:23 +02:00
Louis Dureuil	fdc1763838	Use MultiOps for resolve_query_graph	2023-05-02 18:54:09 +02:00
Louis Dureuil	75819bc940	Remove too many arguments on resolve_maximally_reduced_query_graph	2023-05-02 18:53:40 +02:00
Louis Dureuil	7b8cc25625	rename located_query_terms_from_string -> located_query_terms_from_tokens	2023-05-02 18:53:01 +02:00
meili-bors[bot]	2be641f373	Merge #3718 3718: Fix broken README links r=curquiza a=Kerollmops This PR fixes #3708 by changing the link to the new SDKs and API Reference pages. I would like to thank `@Tommy-42,` who also found the issue. Co-authored-by: Clément Renault <clement@meilisearch.com>	2023-05-02 16:23:38 +00:00
curquiza	ddcb661c19	Use new bors URL of our self hosted instance	2023-05-02 18:20:12 +02:00
Jean-Sébastien Bour	d09b771bce	Add a dedicated error variant for planned failures in index scheduler tests Fixes #3086	2023-05-02 14:37:20 +02:00
Clément Renault	d89d2efb7e	Change a the text of a link	2023-05-02 13:53:36 +02:00
Clément Renault	f284a9c0dd	Fix the README.md broken links	2023-05-02 13:51:50 +02:00
bors[bot]	134e7fc433	Merge #3709 3709: Add SDKs test in a CI r=Kerollmops a=curquiza Add a CI running every week to run the `nightly` docker image of Meilisearch with the most "strategic" SDKs (most used, well tested, strongly typed SDK) - meilisearch-js - instant-meilisearch - meilisearch-php - meilisearch-python - meilisearch-go - meilisearch-ruby - meilisearch-rust Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2023-05-02 11:22:09 +00:00
Clémentine Urquizar	0cba919228	Add SDKs test in a CI	2023-05-02 11:53:28 +02:00
Loïc Lecrenier	aa63091752	Fix bug in exact_attribute	2023-05-02 10:48:32 +02:00
Loïc Lecrenier	58735d6d8f	Fix outdated relevancy test	2023-05-02 10:48:32 +02:00
Loïc Lecrenier	1b514517f5	Fix bug in computation of query term at a position	2023-05-02 10:48:32 +02:00
Loïc Lecrenier	11f814821d	Minor cleanup	2023-05-02 10:48:32 +02:00
Loïc Lecrenier	30fb1153cc	Speed up graph based ranking rule when a lot of different costs exist	2023-05-02 09:59:42 +02:00
Loïc Lecrenier	3b2c8b9f25	Improve performance of position rr	2023-05-02 09:59:42 +02:00
Loïc Lecrenier	2a7f9adf78	Build query graph more correctly from paths Update snapshots	2023-05-02 09:59:42 +02:00
Loïc Lecrenier	608ceea440	Fix bug in position rr	2023-05-02 09:59:42 +02:00
Loïc Lecrenier	79001b9c97	Improve performance of the cheapest path finder algorithm	2023-05-02 09:59:42 +02:00
Loïc Lecrenier	59b12fca87	Fix errors, clippy warnings, and add review comments	2023-04-29 11:48:11 +02:00
Loïc Lecrenier	48f5bb1693	Implements the geo-sort ranking rule	2023-04-29 11:02:16 +02:00
Loïc Lecrenier	93188b3c88	Fix indexing of word_prefix_fid_docids	2023-04-29 10:56:48 +02:00
Loïc Lecrenier	bc4efca611	Add more tests for the attribute ranking rule	2023-04-29 10:56:48 +02:00
TATHAGATA ROY	feaf25a95d	Updated messages pointing to the docs website	2023-04-28 20:52:03 +00:00
bors[bot]	414b3fae89	Merge #3571 3571: Introduce two filters to select documents with `null` and empty fields r=irevoire a=Kerollmops # Pull Request ## Related issue This PR implements the `X IS NULL`, `X IS NOT NULL`, `X IS EMPTY`, `X IS NOT EMPTY` filters that [this comment](https://github.com/meilisearch/product/discussions/539#discussioncomment-5115884) is describing in a very detailed manner. ## What does this PR do? ### `IS NULL` and `IS NOT NULL` This PR will be exposed as a prototype for now. Below is the copy/pasted version of a spec that defines this filter. - `IS NULL` matches fields that `EXISTS` AND `= IS NULL` - `IS NOT NULL` matches fields that `NOT EXISTS` OR `!= IS NULL` 1. `{"name": "A", "price": null}` 2. `{"name": "A", "price": 10}` 3. `{"name": "A"}` `price IS NULL` would match 1 `price IS NOT NULL` or `NOT price IS NULL` would match 2,3 `price EXISTS` would match 1, 2 `price NOT EXISTS` or `NOT price EXISTS` would match 3 common query : `(price EXISTS) AND (price IS NOT NULL)` would match 2 ### `IS EMPTY` and `IS NOT EMPTY` - `IS EMPTY` matches Array `[]`, Object `{}`, or String `""` fields that `EXISTS` and are empty - `IS NOT EMPTY` matches fields that `NOT EXISTS` OR are not empty. 1. `{"name": "A", "tags": null}` 2. `{"name": "A", "tags": [null]}` 3. `{"name": "A", "tags": []}` 4. `{"name": "A", "tags": ["hello","world"]}` 5. `{"name": "A", "tags": [""]}` 6. `{"name": "A"}` 7. `{"name": "A", "tags": {}}` 8. `{"name": "A", "tags": {"t1":"v1"}}` 9. `{"name": "A", "tags": {"t1":""}}` 10. `{"name": "A", "tags": ""}` `tags IS EMPTY` would match 3,7,10 `tags IS NOT EMPTY` or `NOT tags IS EMPTY` would match 1,2,4,5,6,8,9 `tags IS NULL` would match 1 `tags IS NOT NULL` or `NOT tags IS NULL` would match 2,3,4,5,6,7,8,9,10 `tags EXISTS` would match 1,2,3,4,5,7,8,9,10 `tags NOT EXISTS` or `NOT tags EXISTS` would match 6 common query : `(tags EXISTS) AND (tags IS NOT NULL) AND (tags IS NOT EMPTY)` would match 2,4,5,8,9 ## What should the reviewer do? - Check that I tested the filters - Check that I deleted the ids of the documents when deleting documents Co-authored-by: Clément Renault <clement@meilisearch.com> Co-authored-by: Kerollmops <clement@meilisearch.com>	2023-04-27 13:14:00 +00:00
Loïc Lecrenier	899baa0ea5	Update forgotten snapshot from previous commit	2023-04-27 13:43:04 +02:00
Loïc Lecrenier	374095d42c	Add tests for stop words and fix a couple of bugs	2023-04-27 13:30:09 +02:00
Loïc Lecrenier	dd007dceca	Merge pull request #3703 from meilisearch/search-refactor-test-typo-tolerance Search refactor test typo tolerance + some bugfixes	2023-04-27 11:01:35 +02:00
bors[bot]	3ae587205c	Merge #3464 3464: Remove CLI changes for clippy r=curquiza a=dureuill # Pull Request ## Related issue Reverts #3434, which was linked to https://github.com/rust-lang/rust-clippy/issues/10087, as putting the lint in the pedantic group [is being uplifted to Rust 1.67.1](https://github.com/rust-lang/rust/pull/107743#issue-1573438821) (my thanks to everyone involved in this work 🎉). ## Motivation - Using "standard issue" clippy in the CI spares our contributors and us from knowing/remembering to add the lint when running clippy locally - In particular, spares us from configuring tools like rust-analyzer to take the lint into account. - Should this lint come back in another form in the future, we won't blindly ignore it, and we will be able to reassess it, which will be good wrt writing idiomatic Rust. By the time this occurs, lints might be configurable through `clippy.toml` too, which would make disabling one globally much more convenient if needs be. ## Note We should wait for the release of Rust 1.67.1 and its propagation to our CI before merging this. The PR won't pass CI before this. Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-04-26 17:36:56 +00:00
ManyTheFish	1bf2694604	Update cargo lock	2023-04-26 17:41:29 +02:00
Louis Dureuil	ed9cc1af55	Remove CLI changes for clippy	2023-04-26 17:04:09 +02:00
Louis Dureuil	b41a6cbd7a	Check sort criteria also in placeholder search	2023-04-26 16:28:17 +02:00
Louis Dureuil	c8af572697	Add tests for exact words and exact attributes	2023-04-26 16:13:01 +02:00
ManyTheFish	249053e514	Update feature flags	2023-04-26 14:59:25 +02:00
ManyTheFish	ff2cf2a5ae	Update charabia in milli	2023-04-26 14:56:54 +02:00
Loïc Lecrenier	b448aca49c	Add more tests for exactness rr	2023-04-26 11:04:18 +02:00
Loïc Lecrenier	55bad07c16	Fix bug in exact_attribute rr implementation	2023-04-26 10:40:05 +02:00
bors[bot]	380469665f	Merge #3696 3696: Remove the unused snapshot files r=dureuill a=irevoire While « reverting by hand » the PR about the auto batching of the addition&deletion of documents, we forgot to remove the associated snapshot files. Here is the command I used to generate this PR: `cargo insta test --delete-unreferenced-snapshots` Co-authored-by: Tamo <tamo@meilisearch.com>	2023-04-26 07:18:29 +00:00
Loïc Lecrenier	3421125a55	Prevent the `exactness` ranking rule from removing random words Make it strictly follow the term matching strategy	2023-04-26 09:09:19 +02:00
Tamo	0b2200e6e7	remove the unused snapshot files	2023-04-25 17:55:27 +02:00
bors[bot]	0fd5ab9fcc	Merge #3695 3695: Update clippy toolchain from v1.67 to v1.69 r=Kerollmops a=curquiza Co-authored-by: curquiza <clementine@meilisearch.com> Co-authored-by: Clément Renault <clement@meilisearch.com>	2023-04-25 15:45:18 +00:00
Clément Renault	14293f6c8f	Make rustfmt happy	2023-04-25 16:55:39 +02:00
Loïc Lecrenier	d3a94e8b25	Fix bugs and add tests to exactness ranking rule	2023-04-25 16:49:08 +02:00
bors[bot]	1944077a7f	Merge #3566 3566: Improve docker cache r=curquiza a=inductor # Pull Request ## Related issue Fixes #<issue_number> ## What does this PR do? - Use `--mount=type=cache` and GHA build cache for faster build - `=> => transferring context: 75.37MB` to `=> => transferring context: 19.21MB` with `.dockerignore` ## PR checklist Please check if your PR fulfills the following requirements: - [ ] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: inductor <kela@inductor.me> Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com>	2023-04-25 14:49:08 +00:00
Clémentine Urquizar - curqui	8195d366fa	Update .dockerignore	2023-04-25 16:48:25 +02:00
Clément Renault	cfd1b2cc97	Fix the clippy warnings	2023-04-25 16:40:32 +02:00
bors[bot]	19b044b4e6	Merge #3694 3694: Remove Uffizzi because not used by the team r=Kerollmops a=curquiza After discussion with the team, we don't really use Uffizzi and even had issues with it recently: the preview build failing randomly leading to unwanted GitHub notifications + issue to reach the container <img width="628" alt="Capture d’écran 2023-04-25 à 11 55 39" src="https://user-images.githubusercontent.com/20380692/234298586-ef9cff85-ded8-4ec5-ba13-bb7a24d476b3.png"> Thanks for the involvement of Uffizzi team anyway, the tool is just not adapted to our team 😊 Co-authored-by: curquiza <clementine@meilisearch.com>	2023-04-25 14:14:16 +00:00
curquiza	e0730b55b3	Update clippy toolchain from v1.67 to v1.69	2023-04-25 16:05:28 +02:00
curquiza	729fa3770d	Remove Uffizzi because not used by the team	2023-04-25 15:50:38 +02:00
bors[bot]	9cbc85b2f9	Merge #3661 3661: Bump the dependencies r=Kerollmops a=Kerollmops This PR bumps all the dependencies of Meilisearch and the sub-crates. I first did a `cargo upgrade --compatible`, fixed the tests, continued with a `cargo upgrade --incompatible`, and finally fixed the compilation issues. I wasn't able to bump _rustls_ to _0.21.0_ (_actix-web_ is using _tokio-tls 0.23.4_, which uses the _0.20.8_) and neither _vergen_, which changed everything without any guide, I didn't find a way to declare that with the version 8.1.1. `bc25f378e8/meilisearch/build.rs (L4-L9)` Fixes #3285 Co-authored-by: Kerollmops <clement@meilisearch.com>	2023-04-25 08:23:01 +00:00
Kerollmops	a3cf104736	Fix the compilation	2023-04-24 17:50:58 +02:00
Kerollmops	a109802d45	Upgrade the incompatible versions of the dependencies	2023-04-24 17:50:57 +02:00
Kerollmops	2d8060df80	Fix the tests	2023-04-24 17:50:57 +02:00
Kerollmops	47b66e49b8	Upgrade the compatible versions of the dependencies	2023-04-24 17:50:52 +02:00
Loïc Lecrenier	8f2e971879	Add tests for "exactness" rr, make correct universe computation	2023-04-24 16:57:34 +02:00
bors[bot]	654a3a9e19	Merge #3688 3688: Following release v1.1.1: bring back changes into `main` r=curquiza a=curquiza `@meilisearch/engine-team` ensure the changes we bring to `main` are the ones you want Co-authored-by: Louis Dureuil <louis@meilisearch.com> Co-authored-by: bors[bot] <26634292+bors[bot]@users.noreply.github.com> Co-authored-by: Tamo <tamo@meilisearch.com> Co-authored-by: dureuill <dureuill@users.noreply.github.com>	2023-04-24 11:38:23 +00:00
Loïc Lecrenier	d1fdbb63da	Make all search tests pass, fix distinctAttribute bug	2023-04-24 12:12:08 +02:00
bors[bot]	fb9d9239b2	Merge #3674 3674: Bump h2 from 0.3.15 to 0.3.17 r=Kerollmops a=dependabot[bot] Bumps [h2](https://github.com/hyperium/h2) from 0.3.15 to 0.3.17. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/hyperium/h2/releases">h2's releases</a>.</em></p> <blockquote> <h2>v0.3.17</h2> <h2>What's Changed</h2> <ul> <li>Add <code>Error::is_library()</code> method to check if the originated inside <code>h2</code>.</li> <li>Add <code>max_pending_accept_reset_streams(usize)</code> option to client and server builders.</li> <li>Fix theoretical memory growth when receiving too many HEADERS and then RST_STREAM frames faster than an application can accept them off the queue. (CVE-2023-26964)</li> </ul> <h2>v0.3.16</h2> <h2>What's Changed</h2> <ul> <li>Set <code>Protocol</code> extension on requests when received Extended CONNECT requests.</li> <li>Remove <code>B: Unpin + 'static</code> bound requiremented of bufs</li> <li>Fix releasing of frames when stream is finished, reducing memory usage.</li> <li>Fix panic when trying to send data and connection window is available, but stream window is not.</li> <li>Fix spurious wakeups when stream capacity is not available.</li> </ul> <h2>New Contributors</h2> <ul> <li><a href="https://github.com/vi"><code>`@vi</code></a>` made their first contribution in <a href="https://redirect.github.com/hyperium/h2/pull/646">hyperium/h2#646</a></li> <li><a href="https://github.com/silence-coding"><code>`@silence-coding</code></a>` made their first contribution in <a href="https://redirect.github.com/hyperium/h2/pull/651">hyperium/h2#651</a></li> <li><a href="https://github.com/gtsiam"><code>`@gtsiam</code></a>` made their first contribution in <a href="https://redirect.github.com/hyperium/h2/pull/649">hyperium/h2#649</a></li> <li><a href="https://github.com/howardjohn"><code>`@howardjohn</code></a>` made their first contribution in <a href="https://redirect.github.com/hyperium/h2/pull/658">hyperium/h2#658</a></li> <li><a href="https://github.com/cloneable"><code>`@cloneable</code></a>` made their first contribution in <a href="https://redirect.github.com/hyperium/h2/pull/655">hyperium/h2#655</a></li> <li><a href="https://github.com/aftersnow"><code>`@aftersnow</code></a>` made their first contribution in <a href="https://redirect.github.com/hyperium/h2/pull/657">hyperium/h2#657</a></li> <li><a href="https://github.com/vadim-eg"><code>`@vadim-eg</code></a>` made their first contribution in <a href="https://redirect.github.com/hyperium/h2/pull/661">hyperium/h2#661</a></li> </ul> </blockquote> </details> <details> <summary>Changelog</summary> <p><em>Sourced from <a href="https://github.com/hyperium/h2/blob/master/CHANGELOG.md">h2's changelog</a>.</em></p> <blockquote> <h1>0.3.17 (April 13, 2023)</h1> <ul> <li>Add <code>Error::is_library()</code> method to check if the originated inside <code>h2</code>.</li> <li>Add <code>max_pending_accept_reset_streams(usize)</code> option to client and server builders.</li> <li>Fix theoretical memory growth when receiving too many HEADERS and then RST_STREAM frames faster than an application can accept them off the queue. (CVE-2023-26964)</li> </ul> <h1>0.3.16 (February 27, 2023)</h1> <ul> <li>Set <code>Protocol</code> extension on requests when received Extended CONNECT requests.</li> <li>Remove <code>B: Unpin + 'static</code> bound requiremented of bufs</li> <li>Fix releasing of frames when stream is finished, reducing memory usage.</li> <li>Fix panic when trying to send data and connection window is available, but stream window is not.</li> <li>Fix spurious wakeups when stream capacity is not available.</li> </ul> </blockquote> </details> <details> <summary>Commits</summary> <ul> <li><a href="`af4bcacf6d`"><code>af4bcac</code></a> v0.3.17</li> <li><a href="`d3f37e9fba`"><code>d3f37e9</code></a> feat: add <code>max_pending_accept_reset_streams(n)</code> options</li> <li><a href="`5bc8e72e5f`"><code>5bc8e72</code></a> fix: limit the amount of pending-accept reset streams</li> <li><a href="`8088ca658d`"><code>8088ca6</code></a> feat: add Error::is_library method</li> <li><a href="`481c31d528`"><code>481c31d</code></a> chore: Use Cargo metadata for the MSRV build job</li> <li><a href="`d3d50ef812`"><code>d3d50ef</code></a> chore: Replace unmaintained/outdated GitHub Actions</li> <li><a href="`45b9bccdfc`"><code>45b9bcc</code></a> chore: set rust-version in Cargo.toml (<a href="https://redirect.github.com/hyperium/h2/issues/664">#664</a>)</li> <li><a href="`b9dcd39915`"><code>b9dcd39</code></a> v0.3.16</li> <li><a href="`96caf4fca3`"><code>96caf4f</code></a> Add a message for EOF-related broken pipe errors (<a href="https://redirect.github.com/hyperium/h2/issues/615">#615</a>)</li> <li><a href="`732319039f`"><code>7323190</code></a> Avoid spurious wakeups when stream capacity is not available (<a href="https://redirect.github.com/hyperium/h2/issues/661">#661</a>)</li> <li>Additional commits viewable in <a href="https://github.com/hyperium/h2/compare/v0.3.15...v0.3.17">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=h2&package-manager=cargo&previous-version=0.3.15&new-version=0.3.17)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) Dependabot will resolve any conflicts with this PR as long as you don't alter it yourself. You can also trigger a rebase manually by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) You can disable automated security fix PRs for this repo from the [Security Alerts page](https://github.com/meilisearch/meilisearch/network/alerts). </details> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2023-04-24 09:34:35 +00:00
Loïc Lecrenier	a7a0891210	Update examples	2023-04-24 10:07:49 +02:00
Loïc Lecrenier	84d9c731f8	Fix bug in encoding of word_position_docids and word_fid_docids	2023-04-24 09:59:30 +02:00
inductor	11f4724957	ignore all .git	2023-04-18 16:32:31 +09:00
inductor	85182497ab	revert mount	2023-04-18 15:15:33 +09:00
inductor	3e4a356638	EOF	2023-04-18 15:14:13 +09:00
inductor	dfd9c384aa	use docker cache	2023-04-18 15:14:13 +09:00
dependabot[bot]	f0b4046c43	Bump h2 from 0.3.15 to 0.3.17 Bumps [h2](https://github.com/hyperium/h2) from 0.3.15 to 0.3.17. - [Release notes](https://github.com/hyperium/h2/releases) - [Changelog](https://github.com/hyperium/h2/blob/master/CHANGELOG.md) - [Commits](https://github.com/hyperium/h2/compare/v0.3.15...v0.3.17) --- updated-dependencies: - dependency-name: h2 dependency-type: indirect ... Signed-off-by: dependabot[bot] <support@github.com>	2023-04-13 17:03:48 +00:00
bors[bot]	4b953d62fb	Merge #3673 3673: Handle the task queue being full r=irevoire a=dureuill # Pull Request ## Related issue Fixes a remaining issue with #3659 where it was not always possible to send tasks back even after deleting some tasks when prompted. ## Tests - see integration test - also manually tested with a 1MiB task queue. Was not possible to become unblocked before this PR, is now possible. ## What does this PR do? - Use the `non_free_pages_size` method to compute the space occupied by the task db instead of the `real_disk_size` which is not always affected by task deletion. - Expand the test so that it adds a task after the deletion. The test now fails before this PR and succeeds after this PR. Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-04-13 16:24:16 +00:00
Louis Dureuil	c2f4b6ced0	Test: await for the deletion task to complete before trying to add another task	2023-04-13 18:22:42 +02:00
Louis Dureuil	1e6cbcaf12	Update test comment Co-authored-by: Tamo <tamo@meilisearch.com>	2023-04-13 17:27:12 +02:00
Louis Dureuil	066c6bd875	test task db full now checks that a task can be successfully added after deleting tasks	2023-04-13 17:20:06 +02:00
Louis Dureuil	fd583501d7	Use non_free_pages_size instead of real_disk_size to check task db space taken	2023-04-13 17:07:44 +02:00
bors[bot]	bff4bde0ce	Merge #3672 3672: Update version for the next release (v1.1.1) in Cargo.toml r=dureuill a=meili-bot ⚠️ This PR is automatically generated. Check the new version is the expected one and Cargo.lock has been updated before merging. Co-authored-by: dureuill <dureuill@users.noreply.github.com>	2023-04-13 13:34:29 +00:00
dureuill	cd45d21d6e	Update version for the next release (v1.1.1) in Cargo.toml	2023-04-13 13:25:10 +00:00
bors[bot]	f9960be115	Merge #3659 3659: stops receiving tasks once the task queue is full r=Kerollmops a=irevoire Give 20GiB to the task queue + once 50% of the task queue is used, it blocks itself and only receives task deletion requests to ensure we never get in a state where we can’t do anything. Also, create a new error message when we reach this case: ``` Meilisearch cannot receive write operations because the size limit of the tasks database has been reached. Please delete tasks to continue performing write operations. ``` Co-authored-by: Tamo <tamo@meilisearch.com>	2023-04-13 09:11:12 +00:00
Loïc Lecrenier	bd9aba4d77	Add "position" part of the attribute ranking rule	2023-04-13 10:46:09 +02:00
Loïc Lecrenier	8edad8291b	Add logger to attribute rr, fix a bug	2023-04-13 10:25:00 +02:00
Tamo	b3f60ee805	try to fix the ci	2023-04-13 10:18:58 +02:00
Loïc Lecrenier	5acf953298	Merge branch 'search-refactor-attribute-ranking-rule' into search-refactor	2023-04-13 08:28:17 +02:00
Kerollmops	d9cebff61c	Add a simple test to check that attributes are ranking correctly	2023-04-13 08:27:09 +02:00
Loïc Lecrenier	30f7bd03f6	Fix compiler warning/errors caused by previous merge	2023-04-13 08:27:09 +02:00
Kerollmops	df0d9bb878	Introduce the attribute ranking rule in the list of ranking rules	2023-04-13 08:27:09 +02:00
Kerollmops	5230ddb3ea	Resolve the attribute ranking rule conditions	2023-04-13 08:27:09 +02:00
Kerollmops	d6a7c28e4d	Implement the attribute ranking rule edge computation	2023-04-13 08:27:09 +02:00
Kerollmops	e55efc419e	Introduce a new cache for the words fids	2023-04-13 08:27:09 +02:00
Loïc Lecrenier	644e136aee	Merge branch 'search-refactor-typo-attributes' into search-refactor	2023-04-13 08:26:56 +02:00
bors[bot]	ec0ecb5515	Merge #3666 3666: Update README to reference new docs website r=curquiza a=guimachiavelli With the launch of the new website, we need to update the README so it references the correct URLs. Two minor details: - we have removed the contact page from the documentation (it had the same links present in this readme and on the community section of the landing page) - we have recently separated filtering and faceted search into two separate articles Co-authored-by: gui machiavelli <hey@guimachiavelli.com>	2023-04-12 17:30:51 +00:00
Tamo	b4fabce36d	update the error message + update the task db size to 20GiB with a limit at 50%	2023-04-12 18:54:11 +02:00
Tamo	9350a7b017	improve the test and try to understand the issue happening on windows	2023-04-12 18:54:11 +02:00
Tamo	be69ab320d	stops receiving tasks once the task queue is full	2023-04-12 18:54:11 +02:00
bors[bot]	d59d75c9cd	Merge #3667 3667: Disable autobatching of additions and deletions r=irevoire a=dureuill # Pull Request ## Related issue Fixes #3664 ## What does this PR do? - Modifies the autobatcher to not batch document additions and deletions, as a workaround to the DB corruption in #3664 Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-04-12 16:51:13 +00:00
Louis Dureuil	38b7b31beb	Decide to use prefix DB if the word is not an ngram	2023-04-12 16:45:38 +02:00
Louis Dureuil	7a01f20df7	Use word_prefix_docids, make get_word_prefix_docids private	2023-04-12 16:45:38 +02:00
Louis Dureuil	c20c38a7fa	Add SearchContext::word_prefix_docids() method	2023-04-12 16:44:43 +02:00
Louis Dureuil	5ab46324c4	Everyone uses the SearchContext::word_docids instead of get_db_word_docids make get_db_word_docids private	2023-04-12 16:44:43 +02:00
Louis Dureuil	325f17488a	Add SearchContext::word_docids() method	2023-04-12 16:37:05 +02:00
Louis Dureuil	e7ff987c46	Update call sites	2023-04-12 16:36:38 +02:00
Louis Dureuil	244003e36f	Refactor DB cache to return Roaring Bitmaps directly instead of byte slices	2023-04-12 16:35:48 +02:00
Loïc Lecrenier	1f813a6f3b	Simplify implementation of the detailed (=visual) logger	2023-04-12 16:32:53 +02:00
Loïc Lecrenier	96183e804a	Simplify the logger	2023-04-12 16:32:53 +02:00
gui machiavelli	5cfb066b0a	Update README.md	2023-04-12 16:29:20 +02:00
gui machiavelli	a5f44a5ceb	Update references to new docs website With the launch of the new website, we need to update the README so it references the correct URLs. Two minor details: - we have removed the contact page from the documentation (it had the same links present in this readme and on the community section of the landing page) - we have recently separated filtering and faceted search into two separate articles	2023-04-12 16:27:04 +02:00
Loïc Lecrenier	7ab48ed8c7	Matching words fixes	2023-04-12 16:21:43 +02:00
Louis Dureuil	a94e78ffb0	Disable autobatching of additions and deletions	2023-04-12 10:53:00 +02:00
Loïc Lecrenier	e7bb8c940f	Merge branch 'search-refactor-highlighter' into search-refactor-highlighter-merged	2023-04-11 12:22:34 +02:00
Loïc Lecrenier	8cb85294ef	Remove unused import warning	2023-04-07 11:09:30 +02:00
Loïc Lecrenier	d0e9d65025	Fix distinct attribute bugs	2023-04-07 11:09:01 +02:00
Loïc Lecrenier	540a396e49	Fix indexing bug in words_prefix_position	2023-04-07 11:08:39 +02:00
Loïc Lecrenier	a81165f0d8	Merge remote-tracking branch 'origin/main' into search-refactor	2023-04-07 10:15:55 +02:00
Loïc Lecrenier	d6585eb10b	Avoid splitting ngrams into their original component words	2023-04-07 10:13:49 +02:00
Loïc Lecrenier	f7d90ad19f	Merge remote-tracking branch 'origin/search-refactor-tests-doc' into search-refactor	2023-04-07 10:13:18 +02:00
bors[bot]	bc25f378e8	Merge #3647 3647: Improve the health route by ensuring lmdb is not down r=irevoire a=irevoire Fixes #3644 In this PR, I try to make a small read on the `AuthController` and `IndexScheduler` databases. The idea is not to validate that everything works but just to avoid the bug we had last time when lmdb was stuck forever. In order to get access to the `AuthController` without going through the extractor, I need to wrap it in the `Data` type from `actix-web`. And to do that, I had to patch our extractor so it works with the `Data` type as well. Co-authored-by: Tamo <tamo@meilisearch.com>	2023-04-06 18:23:52 +00:00
Louis Dureuil	31630c85d0	exactness graph rr: Add important TODO/FIXME after review	2023-04-06 17:50:39 +02:00
Louis Dureuil	ab09dc0167	exact_attributes: Add TODOs and additional check after review	2023-04-06 17:50:39 +02:00
Louis Dureuil	618c54915d	exact_attribute: dedup nodes after sorting them	2023-04-06 17:50:39 +02:00
Loïc Lecrenier	130d2061bd	Fix indexing of word_position_docid and fid	2023-04-06 17:50:39 +02:00
Louis Dureuil	66ddee4390	Fix word_position_docids indexing	2023-04-06 17:50:39 +02:00
Louis Dureuil	90a6c01495	Use correct codec in proximity	2023-04-06 17:50:39 +02:00
Louis Dureuil	e58426109a	Fix panics and issues in exactness graph ranking rule	2023-04-06 17:50:39 +02:00
Louis Dureuil	f513cf930a	Exact attribute with state	2023-04-06 17:50:39 +02:00
Louis Dureuil	8a13ed7e3f	Add exactness ranking rules	2023-04-06 17:50:39 +02:00
Louis Dureuil	1b8e4d0301	Add ExactTerm and helper method	2023-04-06 17:50:39 +02:00
Louis Dureuil	996619b22a	Increase position by 8 on hard separator when building query terms	2023-04-06 17:50:39 +02:00
Louis Dureuil	2c9822a337	Rename `is_multiple_words` to `is_ngram` and `zero_typo` to `exact`	2023-04-06 17:50:39 +02:00
Louis Dureuil	7276deee0a	Add new db caches	2023-04-06 17:50:39 +02:00
bors[bot]	6a068fe36a	Merge #3649 3649: Update the prototype section in CONTRIBUTING.md r=curquiza a=curquiza Following the creation of this guide https://github.com/meilisearch/engine-team/blob/main/resources/prototypes.md and avoid redundant information. Co-authored-by: curquiza <clementine@meilisearch.com>	2023-04-06 15:27:49 +00:00
ManyTheFish	f7e7f438f8	Patch prefix match	2023-04-06 17:22:31 +02:00
curquiza	8d826e478f	Update the prototype section in CONTRIBUTING.md	2023-04-06 17:10:00 +02:00
ManyTheFish	ba8dcc2d78	Fix clippy	2023-04-06 15:50:47 +02:00
Tamo	4d308d5237	Improve the health route by ensuring lmdb is not down And refactorize slightly the auth controller.	2023-04-06 15:31:42 +02:00
Loïc Lecrenier	7ca91ebb71	Merge branch 'search-refactor-exactness' into search-refactor-tests-doc	2023-04-06 15:16:35 +02:00
ManyTheFish	1ba8a40d61	Remove formating benchmark because they can't be isoloated easily anymore	2023-04-06 15:10:16 +02:00
ManyTheFish	47f6a3ad3d	Take into account that a logger need the search context	2023-04-06 15:02:23 +02:00
bors[bot]	b4c01581cd	Merge #3641 3641: Bring back changes from `release v1.1.0` into `main` after v1.1.0 release r=curquiza a=curquiza Replace https://github.com/meilisearch/meilisearch/pull/3637 since we don't want to pull commits from `main` into `release-v1.1.0` when fixing git conflicts Co-authored-by: ManyTheFish <many@meilisearch.com> Co-authored-by: bors[bot] <26634292+bors[bot]@users.noreply.github.com> Co-authored-by: Charlotte Vermandel <charlottevermandel@gmail.com> Co-authored-by: Tamo <tamo@meilisearch.com> Co-authored-by: Louis Dureuil <louis@meilisearch.com> Co-authored-by: curquiza <clementine@meilisearch.com> Co-authored-by: Clément Renault <clement@meilisearch.com> Co-authored-by: Many the fish <many@meilisearch.com>	2023-04-06 12:37:54 +00:00
ManyTheFish	ae17c62e24	Remove warnings	2023-04-06 14:07:18 +02:00
ManyTheFish	a1148c09c2	remove old matcher	2023-04-06 14:00:21 +02:00
ManyTheFish	9c5f64769a	Integrate the new Highlighter in the search	2023-04-06 13:58:56 +02:00
ManyTheFish	ebe23b04c9	Make the matcher consume the search context	2023-04-06 12:28:28 +02:00
ManyTheFish	13b7c826c1	add new highlighter	2023-04-06 12:15:37 +02:00
Tamo	67fd3b08ef	wait until all tasks are processed before running our dump integration tests	2023-04-05 18:35:43 +02:00
Loïc Lecrenier	5440f43fd3	Fix indexing of word_position_docid and fid	2023-04-05 18:14:00 +02:00
Louis Dureuil	d9460a76f4	Fix word_position_docids indexing	2023-04-05 18:14:00 +02:00
Louis Dureuil	d1ddaa223d	Use correct codec in proximity	2023-04-05 18:14:00 +02:00
Louis Dureuil	f7ecea142e	Fix panics and issues in exactness graph ranking rule	2023-04-05 18:13:46 +02:00
Louis Dureuil	337e75b0e4	Exact attribute with state	2023-04-05 18:12:46 +02:00
Loïc Lecrenier	b5691802a3	Add new tests and fix construction of query graph from paths	2023-04-05 16:31:10 +02:00
bors[bot]	1690aec7f1	Merge #3638 3638: filter errors - `geo(x,y,z)` and `geoDistance(x,y,z)` r=irevoire a=cymruu # Pull Request ## Related issue Fixes #3006 ## What does this PR do? - fixes the display function of `ParseError::ReservedGeo`. The previous display string was missing back ticks around available filters. - makes the filter-parser parse `_geo(x,y,z)` and `geoDistance(x,y,z)` filters. Both parsing functions will throw an error if the filter was used. - removes `FilterError::ReservedGeo` and `FilterError::Reserved` error variants since they are now thrown by the filter-parser. I ran `cargo test --package milli -- --test-threads 1` and the tests passed. ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [ ] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Filip Bachul <filipbachul@gmail.com> Co-authored-by: filip <filipbachul@gmail.com>	2023-04-05 13:23:50 +00:00
filip	f267bed352	remove a unnecessary comment Co-authored-by: Tamo <irevoire@protonmail.ch>	2023-04-05 13:44:55 +02:00
Loïc Lecrenier	6e50f23896	Add more search tests	2023-04-05 13:33:23 +02:00
Tamo	597d57bf1d	Merge branch 'main' into bring-back-changes-v1.1.0	2023-04-05 11:32:14 +02:00
Loïc Lecrenier	4c8a0179ba	Add more search tests	2023-04-05 11:30:49 +02:00
Loïc Lecrenier	c69cbec64a	Add more search tests	2023-04-05 11:20:04 +02:00
bors[bot]	01ac8344ad	Merge #3643 3643: Add sprint issue to the template issues r=curquiza a=curquiza Following our [internal guide](https://www.notion.so/meilisearch/Delivery-synchronization-policy-b4ec15c3e56a49539fa02d2f1d381de3) presentation, I put the template here Co-authored-by: curquiza <clementine@meilisearch.com>	2023-04-05 08:05:21 +00:00
curquiza	3508ba2f20	Add sprint issue to the template issues	2023-04-04 18:58:43 +02:00
Loïc Lecrenier	ce328c329d	Move bucket sort function to its own module and fix a bug	2023-04-04 18:03:08 +02:00
Loïc Lecrenier	959e4607bb	Add more search tests	2023-04-04 18:02:46 +02:00
Louis Dureuil	4b4ffb8ec9	Add exactness ranking rules	2023-04-04 17:12:07 +02:00
Louis Dureuil	3951fe22ab	Add ExactTerm and helper method	2023-04-04 17:09:32 +02:00
Louis Dureuil	4d5bc9df4c	Increase position by 8 on hard separator when building query terms	2023-04-04 17:07:26 +02:00
Louis Dureuil	ec2f8e8040	Rename `is_multiple_words` to `is_ngram` and `zero_typo` to `exact`	2023-04-04 17:06:07 +02:00
Louis Dureuil	406b8bd248	Add new db caches	2023-04-04 17:04:46 +02:00
Loïc Lecrenier	62b9c6fbee	Add search tests	2023-04-04 16:18:22 +02:00
Loïc Lecrenier	b439d36807	Split query_term module into multiple submodules	2023-04-04 15:38:30 +02:00
Loïc Lecrenier	faceb661e3	Add note that a part of the code needs fixing	2023-04-04 15:02:01 +02:00
Loïc Lecrenier	4129d657e2	Simplify query_term module a bit	2023-04-04 15:01:42 +02:00
Filip Bachul	1e6fe71a67	fix clippy warning	2023-04-03 20:18:26 +02:00
Filip Bachul	0fba08cd72	fmt	2023-04-03 20:18:26 +02:00
Filip Bachul	189d4c3b70	add geoPoint integration tests	2023-04-03 20:18:26 +02:00
Filip Bachul	2fff6f7f23	add parse_geo_distance to parse_primary	2023-04-03 20:18:26 +02:00
Filip Bachul	fddfb37f1f	remove unnecessary FilterError:ReservedGeo and FilterError:ReservedGeo	2023-04-03 20:18:26 +02:00
Filip Bachul	52b4090286	update integration tests	2023-04-03 20:18:26 +02:00
Filip Bachul	3cabfb448b	fix backticks in ErrorKind::ReservedGeo display	2023-04-03 20:18:26 +02:00
Filip Bachul	77cf5b3787	handle _geoDistance(x,y,z) filter error	2023-04-03 20:18:26 +02:00
Filip Bachul	3acc5bbb15	handle _geo(x,y,z) filter error	2023-04-03 20:18:26 +02:00
bors[bot]	114436926f	Merge #3631 3631: sort errors - `_geo(x,y)` and `_geoDistance(x,y)` return `ReservedKeyword` r=irevoire a=cymruu # Pull Request ## Related issue Fix part of #3006 (sort errors) ## What does this PR do? - made meilisearch return `ReservedKeyword` when sorting by `_geo(x,y)` or `_geoDistance(x,y)` Screenshot: ![image](https://user-images.githubusercontent.com/2981598/228969970-56dae3d2-7851-48ea-a913-c4483224d709.png) I ran `cargo test --package milli -- --test-threads 1` and the tests passed. ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [ ] Have you made sure that the title is accurate and descriptive of the changes? Co-authored-by: Filip Bachul <filipbachul@gmail.com>	2023-04-03 16:28:15 +00:00
bors[bot]	0f7904fb38	Merge #3636 3636: Add a newline after the meilisearch version in the issue template r=curquiza a=bidoubiwa # Pull Request ## What does this PR do? Just a small change that makes it easier to remove the version example and adds consistency with the other examples in its positioning. Co-authored-by: cvermand <33010418+bidoubiwa@users.noreply.github.com>	2023-04-03 13:52:08 +00:00
Loïc Lecrenier	3f13608002	Fix computation of ngram derivations	2023-04-03 15:27:49 +02:00
cvermand	590b1d8fb7	Add a newline after the meilisearch version in the issue template Just a small change to make it easier in removing the version example and is consistent with the other examples in its positioning.	2023-04-03 13:14:20 +02:00
Loïc Lecrenier	4708d9b016	Fix compiler warnings/errors	2023-04-03 10:09:27 +02:00
Clément Renault	0d2e7bcc13	Implement the previous way for the exhaustive distinct candidates	2023-04-03 10:08:10 +02:00
Loïc Lecrenier	55fbfb6124	Merge branch 'search-refactor-located-query-terms' into search-refactor	2023-04-03 10:04:36 +02:00
bors[bot]	be9741eb8a	Merge #3633 3633: Bump Swatinem/rust-cache from 2.2.0 to 2.2.1 r=curquiza a=dependabot[bot] Bumps [Swatinem/rust-cache](https://github.com/Swatinem/rust-cache) from 2.2.0 to 2.2.1. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/Swatinem/rust-cache/releases">Swatinem/rust-cache's releases</a>.</em></p> <blockquote> <h2>v2.2.1</h2> <ul> <li>Update <code>`@actions/cache</code>` dependency to fix usage of <code>zstd</code> compression.</li> </ul> </blockquote> </details> <details> <summary>Changelog</summary> <p><em>Sourced from <a href="https://github.com/Swatinem/rust-cache/blob/master/CHANGELOG.md">Swatinem/rust-cache's changelog</a>.</em></p> <blockquote> <h2>2.2.1</h2> <ul> <li>Update <code>`@actions/cache</code>` dependency to fix usage of <code>zstd</code> compression.</li> </ul> <h2>2.2.0</h2> <ul> <li>Add new <code>save-if</code> option to always restore, but only conditionally save the cache.</li> </ul> <h2>2.1.0</h2> <ul> <li>Only hash <code>Cargo.{lock,toml}</code> files in the configured workspace directories.</li> </ul> <h2>2.0.2</h2> <ul> <li>Avoid calling <code>cargo metadata</code> on pre-cleanup.</li> <li>Added <code>prefix-key</code>, <code>cache-directories</code> and <code>cache-targets</code> options.</li> </ul> <h2>2.0.1</h2> <ul> <li>Primarily just updating dependencies to fix GitHub deprecation notices.</li> </ul> <h2>2.0.0</h2> <ul> <li>The action code was refactored to allow for caching multiple workspaces and different <code>target</code> directory layouts.</li> <li>The <code>working-directory</code> and <code>target-dir</code> input options were replaced by a single <code>workspaces</code> option that has the form of <code>$workspace -> $target</code>.</li> <li>Support for considering <code>env-vars</code> as part of the cache key.</li> <li>The <code>sharedKey</code> input option was renamed to <code>shared-key</code> for consistency.</li> </ul> <h2>1.4.0</h2> <ul> <li>Clean both <code>debug</code> and <code>release</code> target directories.</li> </ul> <h2>1.3.0</h2> <ul> <li>Use Rust toolchain file as additional cache key.</li> <li>Allow for a configurable target-dir.</li> </ul> <h2>1.2.0</h2> <ul> <li>Cache <code>~/.cargo/bin</code>.</li> <li>Support for custom <code>$CARGO_HOME</code>.</li> <li>Add a <code>cache-hit</code> output.</li> <li>Add a new <code>sharedKey</code> option that overrides the automatic job-name based key.</li> </ul> <h2>1.1.0</h2> <ul> <li>Add a new <code>working-directory</code> input.</li> <li>Support caching git dependencies.</li> </ul> <!-- raw HTML omitted --> </blockquote> <p>... (truncated)</p> </details> <details> <summary>Commits</summary> <ul> <li><a href="`6fd3edff69`"><code>6fd3edf</code></a> 2.2.1</li> <li><a href="`a1c019f71a`"><code>a1c019f</code></a> update dependencies and rebuild</li> <li><a href="`664ce0090f`"><code>664ce00</code></a> chore: Create check-dist.yml (<a href="https://redirect.github.com/Swatinem/rust-cache/issues/96">#96</a>)</li> <li>See full diff in <a href="https://github.com/Swatinem/rust-cache/compare/v2.2.0...v2.2.1">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=Swatinem/rust-cache&package-manager=github_actions&previous-version=2.2.0&new-version=2.2.1)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) You can trigger a rebase of this PR by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) </details> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2023-04-03 07:53:07 +00:00
Loïc Lecrenier	58fe260c72	Allow removing all the terms from a query if it contains a phrase	2023-04-03 09:18:02 +02:00
Loïc Lecrenier	24e5f6f7a9	Don't remove phrases with "last" term matching strategy	2023-04-03 09:17:33 +02:00
dependabot[bot]	0177d66149	Bump Swatinem/rust-cache from 2.2.0 to 2.2.1 Bumps [Swatinem/rust-cache](https://github.com/Swatinem/rust-cache) from 2.2.0 to 2.2.1. - [Release notes](https://github.com/Swatinem/rust-cache/releases) - [Changelog](https://github.com/Swatinem/rust-cache/blob/master/CHANGELOG.md) - [Commits](https://github.com/Swatinem/rust-cache/compare/v2.2.0...v2.2.1) --- updated-dependencies: - dependency-name: Swatinem/rust-cache dependency-type: direct:production update-type: version-update:semver-patch ... Signed-off-by: dependabot[bot] <support@github.com>	2023-04-01 17:58:46 +00:00
Louis Dureuil	9b87c36200	Limit the number of derivations for a single word.	2023-03-31 09:19:18 +02:00
Filip Bachul	1861c69964	fmt	2023-03-30 23:37:26 +02:00
Filip Bachul	cb2b5eb38e	handle _geoDistance(x,x) sort error	2023-03-30 23:21:23 +02:00
Filip Bachul	53aa0a1b54	handle _geo(x,x) sort error	2023-03-30 23:17:34 +02:00
Loïc Lecrenier	12b26cd54e	Don't remove phrases from the query with term matching strategy Last	2023-03-30 14:54:08 +02:00
Loïc Lecrenier	061b1e6d7c	Tiny refactor of query graph remove_nodes method	2023-03-30 14:49:25 +02:00
Loïc Lecrenier	0d6e8b5c31	Fix phrase search bug when the phrase has only one word	2023-03-30 14:48:12 +02:00
Loïc Lecrenier	d48cdc67a0	Fix term matching strategy bugs	2023-03-30 14:01:52 +02:00
Loïc Lecrenier	35c16ad047	Use new term matching strategy logic in words ranking rule	2023-03-30 13:15:43 +02:00
Loïc Lecrenier	2997d1f186	Use new term matching strategy logic in resolve_maximally_reduced_...	2023-03-30 13:12:51 +02:00
Loïc Lecrenier	2a5997fb20	Avoid expensive assert! in bucket sort function	2023-03-30 13:07:17 +02:00
Loïc Lecrenier	ee8a9e0bad	Remove outdated sentence in documentation	2023-03-30 12:22:24 +02:00
Loïc Lecrenier	3b0737a092	Fix detailed logger	2023-03-30 12:20:44 +02:00
Loïc Lecrenier	fdd02105ac	Graph-based ranking rule + term matching strategy support	2023-03-30 12:19:21 +02:00
Loïc Lecrenier	aa9592455c	Refactor the paths_of_cost algorithm Support conditions that require certain nodes to be skipped	2023-03-30 12:11:11 +02:00
Loïc Lecrenier	01e24dd630	Rewrite proximity ranking rule	2023-03-30 11:59:06 +02:00
Loïc Lecrenier	ae6bb1ce17	Update the ConditionDocidsCache after change to RankingRuleGraphTrait	2023-03-30 11:41:20 +02:00
Loïc Lecrenier	5fd28620cd	Build ranking rule graph correctly after changes to trait definition	2023-03-30 11:32:55 +02:00
Loïc Lecrenier	728710d63a	Update typo ranking rule to use new query term structure	2023-03-30 11:32:19 +02:00
Loïc Lecrenier	fa81381865	Update the trait requirements of ranking-rule graphs	2023-03-30 11:19:45 +02:00
Loïc Lecrenier	b96a682f16	Update resolve_graph module to work with lazy query terms	2023-03-30 11:10:38 +02:00
Loïc Lecrenier	d0f048c068	Simplify the API of the DatabaseCache	2023-03-30 11:08:17 +02:00
Loïc Lecrenier	223e82a10d	Update QueryGraph to use new lazy query terms + build from paths	2023-03-30 11:06:02 +02:00
Loïc Lecrenier	9507ff5e31	Update query term structure to allow for laziness	2023-03-30 11:06:02 +02:00
Louis Dureuil	c2b025946a	`located_query_terms_from_string`: use u16 for positions, hard limit number of iterated tokens. - Refactor phrase logic to reduce number of possible states	2023-03-30 11:04:14 +02:00
bors[bot]	950f73b8bb	Merge #3623 3623: Update mini-dashboard to version v0.2.7 r=curquiza a=bidoubiwa ## Changes * Retrieve the API Key from the url parameters (#416) `@qdequele` ## 🐛 Bug Fixes * Fix show more button not displaying all fields (#419) `@bidoubiwa` Thanks again to `@bidoubiwa,` and `@qdequele!` 🎉 Co-authored-by: Charlotte Vermandel <charlottevermandel@gmail.com>	2023-03-30 08:31:29 +00:00
Loïc Lecrenier	3a818c5e87	Add more functionality to interners	2023-03-30 09:56:23 +02:00
bors[bot]	7871d12025	Merge #3624 3624: Reduce the time to import a dump r=irevoire a=irevoire When importing a dump, this PR does multiple things; - Stops committing the changes between each task import - Stop deserializing + serializing every bitmap for every task Pros: Importing 1M tasks in a dump went from 3m36 on my computer to 6s Cons: We use slightly more memory, but since we’re using roaring bitmaps, that really shouldn’t be noticeable. Fixes #3620 Co-authored-by: Tamo <tamo@meilisearch.com>	2023-03-29 13:40:25 +00:00
Louis Dureuil	d74134ce3a	Check sort criteria	2023-03-29 15:21:54 +02:00
Louis Dureuil	5ac129bfa1	Mark geosearch as currently unimplemented for sort rule	2023-03-29 15:20:42 +02:00
Charlotte Vermandel	e7153e0a97	Update mini-dashboard to version V0.2.7	2023-03-29 14:49:39 +02:00
bors[bot]	37a24a4a05	Merge #3621 3621: Fix facet normalization r=Kerollmops a=ManyTheFish # Pull Request Make sure the facet normalization is the same between indexing and search. ## Related issue Fixes #3599 Co-authored-by: ManyTheFish <many@meilisearch.com>	2023-03-29 12:47:20 +00:00
Tamo	3fb67f94f7	Reduce the time to import a dump by caching some datas With this commit, for a dump containing 1M tasks we went form 1m02 to 6s	2023-03-29 14:44:15 +02:00
ManyTheFish	6592746337	Fix other unrelated tests	2023-03-29 14:36:17 +02:00
Tamo	cf5145b542	Reduce the time to import a dump With this commit, for a dump containing 1M tasks we went from 3m36s to import the task queue down to 1m02s	2023-03-29 14:27:40 +02:00
ManyTheFish	efea1e5837	Fix facet normalization	2023-03-29 12:02:24 +02:00
ManyTheFish	b744f33530	Add test	2023-03-29 12:01:52 +02:00
bors[bot]	31bb61ba99	Merge #3608 3608: In a settings update, check to see if the primary key actually changes before erroring out r=irevoire a=GregoryConrad Previously, if the primary key was set and a Settings update contained a primary key, an error would be returned. However, this error is not needed if the new PK == the current PK. This PR just checks to see if the PK actually changes before raising an error. I came across this slight hiccup in https://github.com/GregoryConrad/mimir/issues/156#issuecomment-1484128654 Co-authored-by: Gregory Conrad <gregorysconrad@gmail.com>	2023-03-29 09:07:51 +00:00
Louis Dureuil	abb4522f76	Small comment on ignored rules for placeholder search	2023-03-29 09:11:06 +02:00
bors[bot]	d4f54fc55e	Merge #3617 3617: update the geoBoundingBox feature r=dureuill a=irevoire Closing #3616 Implementing this change in the spec: `38a715c072` Now instead of using the (top_left, bottom_right) corners of the bounding box, it’s using the (top_right, bottom_left) corners. Co-authored-by: Tamo <tamo@meilisearch.com>	2023-03-29 07:01:17 +00:00
Louis Dureuil	ef084ef042	SmallBitmap: Consistently panic on incoherent universe lengths	2023-03-29 08:45:38 +02:00
Louis Dureuil	3524bd1257	SmallBitmap: Add documentation	2023-03-29 08:44:11 +02:00
Tamo	a50b058557	update the geoBoundingBox feature Now instead of using the (top_left, bottom_right) corners of the bounding box it s using the (top_right, bottom_left) corners.	2023-03-28 18:26:18 +02:00
Louis Dureuil	d4f6216966	Resolve rule time sort criteria	2023-03-28 16:42:02 +02:00
Louis Dureuil	77acafe534	Resolve search time sort criteria for placeholder search	2023-03-28 16:41:03 +02:00
Louis Dureuil	53afda3237	Update search usage in example	2023-03-28 16:35:46 +02:00
Louis Dureuil	abb19d368d	Initialize query time ranking rule for query search	2023-03-28 12:40:52 +02:00
Louis Dureuil	b4a52a622e	BoxRankingRule	2023-03-28 12:39:42 +02:00
Louis Dureuil	8d7d8cdc2f	Clean-up index example	2023-03-27 18:34:10 +02:00
Louis Dureuil	626a93b348	Search example: panic when missing the index path	2023-03-27 18:18:01 +02:00
Louis Dureuil	af65fe201a	Clean-up search example	2023-03-27 17:49:43 +02:00
Louis Dureuil	9b83b1deb0	Expose SearchLogger trait	2023-03-27 17:49:18 +02:00
Louis Dureuil	e9eb271499	Remove empty attribute_rule mod	2023-03-27 11:08:03 +02:00
Louis Dureuil	3281a88d08	SmallBitmap: don't expose internal items	2023-03-27 11:04:43 +02:00
Louis Dureuil	5a644054ab	Removed unused search impl	2023-03-27 11:04:27 +02:00
Louis Dureuil	16fefd364e	Add TODO notes	2023-03-27 11:04:04 +02:00
Gregory Conrad	e7994cdeb3	feat: check to see if the PK changed before erroring out Previously, if the primary key was set and a Settings update contained a primary key, an error would be returned. However, this error is not needed if the new PK == the current PK. This commit just checks to see if the PK actually changes before raising an error.	2023-03-26 12:18:39 -04:00
Loïc Lecrenier	00bad8c716	Add comments suggesting performance improvements	2023-03-23 10:18:24 +01:00
Loïc Lecrenier	862714a18b	Remove criterion_implementation_strategy param of Search	2023-03-23 09:44:12 +01:00
Loïc Lecrenier	d18ebe4f3a	Remove more warnings	2023-03-23 09:41:18 +01:00
Loïc Lecrenier	7169d85115	Remove old query_tree code and make clippy happy	2023-03-23 09:39:16 +01:00
Loïc Lecrenier	f5f5f03ec0	Remove old criteria code	2023-03-23 09:35:53 +01:00
Loïc Lecrenier	9b2653427d	Split position DB into fid and relative position DB	2023-03-23 09:22:01 +01:00
Loïc Lecrenier	56b7209f26	Make clippy happy	2023-03-23 09:16:17 +01:00
Loïc Lecrenier	9b1f439a91	WIP	2023-03-23 09:12:35 +01:00
Loïc Lecrenier	01c7d2de8f	Add example targets to the milli crate	2023-03-22 14:50:41 +01:00
Loïc Lecrenier	a86aeba411	WIP	2023-03-22 14:43:08 +01:00
bors[bot]	514b60f8c8	Merge #3597 3597: ensure that the task queue is correctly imported r=irevoire a=irevoire ## Related issue Fixes #3596 I updated all the dump's integration tests to ensure that we're effectively able to query the tasks Co-authored-by: Tamo <tamo@meilisearch.com>	2023-03-21 17:31:26 +00:00
Tamo	a2b151e877	ensure that the task queue is correctly imported reduce the size of the snapshots file	2023-03-21 14:41:46 +01:00
Loïc Lecrenier	384fdc2df4	Fix two bugs in proximity ranking rule	2023-03-21 11:43:25 +01:00
Loïc Lecrenier	83e5b4ed0d	Compute edges of proximity graph lazily	2023-03-21 10:44:40 +01:00
Loïc Lecrenier	272cd7ebbd	Small cleanup	2023-03-20 13:39:19 +01:00
Loïc Lecrenier	c63c7377e6	Switch order of MappedInterner generic params	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	9259cdb12e	Update Cargo.lock (was mistakenly changed during rebase)	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	5b50e49522	cargo fmt	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	65474c8de5	Update new sort ranking rule after rebasing	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	fbb1ba3de0	Cargo fmt	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	a59ca28e2c	Add forgotten file	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	825f742000	Simplify graph-based ranking rule impl	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	dd491320e5	Simplify graph-based ranking rule impl	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	c6ff97a220	Rewrite the dead-ends cache to detect more dead-ends	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	49240c367a	Fix bug in cost of typo conditions	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	1e6e624078	Fix bug in SmallBitmap	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	8b4e07e1a3	WIP	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	2853009987	Renaming Edge -> Condition	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	aa59c3bc2c	Replace EdgeCondition with an Option<..> + other code cleanup	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	7b1d8f4c6d	Make PathSet strongly typed	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	a49ddec9df	Prune the query graph after executing a ranking rule	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	05fe856e6e	Merge forward and backward proximity conditions in proximity graph	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	c0cdaf9f53	Fix bug in the proximity ranking rule for queries with ngrams	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	e9cf58d584	Refactor of the Interner	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	31628c5cd4	Merge Phrase and WordDerivations into one structure	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	3004e281d7	Support ngram typos + splitwords and splitwords+synonyms in proximity	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	14e8d0aaa2	Rename lifetime	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	1c58cf8426	Intern ranking rule graph edge conditions as well	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	5155fd2bf1	Reorganise initialisation of ranking rules + rename PathsMap -> PathSet	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	9ec9c204d3	Small code cleanup	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	78b9304d52	Implement distinct attribute	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	0465ba4a05	Intern more values	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	2099991dd1	Continue documenting and cleaning up the code	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	c232cdabf5	Add documentation	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	4e266211bf	Small code reorganisation	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	57fa689131	Cargo fmt	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	10626dddfc	Add a few more optimisations to new search algorithms	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	9051065c22	Apply a few optimisations for graph-based ranking rules	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	e8c76cf7bf	Intern all strings and phrases in the search logic	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	3f1729a17f	Update new search test	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	cab2b6bcda	Fix: computation of initial universe, code organisation	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	c4979a2fda	Fix code visibility issue + unimplemented detail in proximity rule	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	23931f8a4f	Fix small bug in visual logger of search algo	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	aa414565bb	Fix proximity graph edge builder to include all proximities	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	1db152046e	WIP on split words and synonyms support	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	c27ea2677f	Rewrite cheapest path algorithm and empty path cache It is now much simpler and has much better performance.	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	caa1e1b923	Add typo ranking rule to new search impl	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	71f18e4379	Add sort ranking rule to new search impl	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	600e3dd1c5	Remove warnings	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	362eb0de86	Add support for filters	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	998d46ac10	Add support for search offset and limit	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	6c85c0d95e	Fix more bugs + visual empty path cache logging	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	0e1fbbf7c6	Fix bugs in query graph's "remove word" and "cheapest paths" algos	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	6806640ef0	Fix d2 description of paths map	2023-03-20 09:41:56 +01:00
Loïc Lecrenier	173e37584c	Improve the visual/detailed search logger	2023-03-20 09:41:55 +01:00
Loïc Lecrenier	6ba4d5e987	Add a search logger	2023-03-20 09:41:55 +01:00
Loïc Lecrenier	dd12d44134	Support swapped word pairs in new proximity ranking rule impl	2023-03-20 09:41:55 +01:00
Loïc Lecrenier	a61495d660	Update Cargo.toml (commit to be deleted later)	2023-03-20 09:41:55 +01:00
Loïc Lecrenier	c8e251bf24	Remove noise in codebase	2023-03-20 09:41:55 +01:00
Loïc Lecrenier	a938fbde4a	Use a cache when resolving the query graph	2023-03-20 09:41:55 +01:00
Loïc Lecrenier	dcf3f1d18a	Remove EdgeIndex and NodeIndex types, prefer u32 instead	2023-03-20 09:41:55 +01:00
Loïc Lecrenier	66d0c63694	Add some documentation and use bitmaps instead of hashmaps when possible	2023-03-20 09:41:55 +01:00
Loïc Lecrenier	132191360b	Introduce the sort ranking rule working with the new search structures	2023-03-20 09:41:55 +01:00
Loïc Lecrenier	345c99d5bd	Introduce the words ranking rule working with the new search structures	2023-03-20 09:41:55 +01:00
Loïc Lecrenier	89d696c1e3	Introduce the proximity ranking rule as a graph-based ranking rule	2023-03-20 09:41:55 +01:00
Loïc Lecrenier	c645853529	Introduce a generic graph-based ranking rule	2023-03-20 09:41:55 +01:00
Loïc Lecrenier	a70ab8b072	Introduce a function to find the K shortest paths in a graph	2023-03-20 09:41:55 +01:00
Loïc Lecrenier	48aae76b15	Introduce a function to find the docids of a set of paths in a graph	2023-03-20 09:41:55 +01:00
Loïc Lecrenier	23bf572dea	Introduce cache structures used with ranking rule graphs	2023-03-20 09:41:55 +01:00
Loïc Lecrenier	864f6410ed	Introduce a structure to represent a set of graph paths efficiently	2023-03-20 09:41:55 +01:00
Loïc Lecrenier	c9bf6bb2fa	Introduce a structure to implement ranking rules with graph algorithms	2023-03-20 09:41:55 +01:00
Loïc Lecrenier	46249ea901	Implement a function to find a QueryGraph's docids	2023-03-20 09:41:55 +01:00
Loïc Lecrenier	ce0d1e0e13	Introduce a common way to manage the coordination between ranking rules	2023-03-20 09:41:55 +01:00
Loïc Lecrenier	5065d8b0c1	Introduce a DatabaseCache to memorize the addresses of LMDB values	2023-03-20 09:41:55 +01:00
Loïc Lecrenier	a83007c013	Introduce structure to represent search queries as graphs	2023-03-20 09:41:55 +01:00
Loïc Lecrenier	79e0a6dd4e	Introduce a new search module, eventually meant to replace the old one The code here does not compile, because I am merely splitting one giant commit into smaller ones where each commit explains a single file.	2023-03-20 09:41:55 +01:00
Loïc Lecrenier	2d88089129	Remove unused term matching strategies	2023-03-20 09:41:55 +01:00
Loïc Lecrenier	1d937f831b	Temporarily remove codegen-units - 1	2023-03-20 09:41:55 +01:00
Loïc Lecrenier	6c659dc12f	Use MiMalloc in milli tests	2023-03-20 09:41:37 +01:00
Clément Renault	a8531053a0	Make sure the parser reject invalid syntax	2023-03-16 11:09:20 +01:00
Clément Renault	cf34d1c95f	Fix a test that forget to match a Null value	2023-03-15 17:17:19 +01:00
Clément Renault	1a9c58a7ab	Fix a bug with the new flattening rules	2023-03-15 16:56:44 +01:00
Clément Renault	64571c8288	Improve the testing of the filters	2023-03-15 14:57:17 +01:00
Clément Renault	72123c458b	Fix the tests to make flattening work	2023-03-15 14:12:34 +01:00
Clément Renault	d5881519cb	Make the json flattener return the original values	2023-03-15 14:12:34 +01:00
Clément Renault	ea016d97af	Implementing an IS EMPTY filter	2023-03-15 14:12:34 +01:00
bors[bot]	70c906d4b4	Merge #3576 3576: Add boolean support for csv documents r=irevoire a=irevoire Fixes https://github.com/meilisearch/meilisearch/issues/3572 ## What does this PR do? Add support for the boolean types in csv documents. The type definition is `boolean` and the possible values are - `true` for true - `false` for false - ` ` for null Here is an example: ```csv #id,cute:boolean 0,true 1,false 2, ``` Co-authored-by: Tamo <tamo@meilisearch.com>	2023-03-14 12:28:12 +00:00
Clément Renault	fa2ea4a379	Update the test to accept the new IS syntax	2023-03-14 10:31:27 +01:00
Clément Renault	030263caa3	Change the IS NULL filter syntax to use the IS keyword	2023-03-14 10:31:04 +01:00
Kerollmops	c25779afba	Specify that the NULL keyword is a keyword too	2023-03-13 17:40:34 +01:00
Tamo	0f33a65468	makes kero happy	2023-03-13 16:51:11 +01:00
bors[bot]	7c9a8b1e1b	Merge #3587 3587: Enable cache again in test suite CI r=curquiza a=curquiza Following the change in this PR introduced in v1.1: https://github.com/meilisearch/meilisearch/pull/3422 The cache was removed due to failures (lack of space). Now the binary is smaller (from 250Mb to 50Mb) we want to try to enable the cache again. Indeed, without the cache step, the CIs are wayyyy slower (45min instead of 20-30min). For later: Rust 1.68 introduced a new way to fetch crates. Updating the rust version might also help in the future! Co-authored-by: curquiza <clementine@meilisearch.com>	2023-03-13 13:51:32 +00:00
curquiza	f45daf8031	Enable cache again in test suite CI	2023-03-13 14:24:15 +01:00
bors[bot]	fb1260ee88	Merge #3568 #3569 3568: CI: Fix `publish-aarch64` job that still uses ubuntu-18.04 r=Kerollmops a=curquiza Fixes #3563 Main change - add the usage of the `ubuntu-18.04` container instead of the native `ubuntu-18.04` of GitHub actions: I had to install docker in the container. Small additional changes - remove useless `fail-fast` and unused/irrelevant matrix inputs (`build`, `linker`, `os`, `use-cross`...) - Remove useless step in job Proof of work with this CI triggered on this current branch: https://github.com/meilisearch/meilisearch/actions/runs/4366233882 3569: Enhance Japanese language detection r=dureuill a=ManyTheFish # Pull Request This PR is a prototype and can be tested by downloading [the dedicated docker image](https://hub.docker.com/layers/getmeili/meilisearch/prototype-better-language-detection-0/images/sha256-a12847de00e21a71ab797879fd09777dadcb0881f65b5f810e7d1ed434d116ef?context=explore): ```bash $ docker pull getmeili/meilisearch:prototype-better-language-detection-0 ``` ## Context Some Languages are harder to detect than others, this miss-detection leads to bad tokenization making some words or even documents completely unsearchable. Japanese is the main Language affected and can be detected as Chinese which has a completely different way of tokenization. A [first iteration has been implemented for v1.1.0](https://github.com/meilisearch/meilisearch/pull/3347) but is an insufficient enhancement to make Japanese work. This first implementation was detecting the Language during the indexing to avoid bad detections during the search. Unfortunately, some documents (shorter ones) can be wrongly detected as Chinese running bad tokenization for these documents and making possible the detection of Chinese during the search because it has been detected during the indexing. For instance, a Japanese document `{"id": 1, "name": "東京スカパラダイスオーケストラ"}` is detected as Japanese during indexing, during the search the query `東京` will be detected as Japanese because only Japanese documents have been detected during indexing despite the fact that v1.0.2 would detect it as Chinese. However if in the dataset there is at least one document containing a field with only Kanjis like: _A document with only 1 field containing only Kanjis:_ ```json { "id":4, "name": "東京特許許可局" } ``` _A document with 1 field containing only Kanjis and 1 field containing several Japanese characters:_ ```json { "id":105, "name": "東京特許許可局", "desc": "日経平均株価は26日に約8カ月ぶりに2万4000円の心理的な節目を上回った。株高を支える材料のひとつは、自民党総裁選で3選を決めた安倍晋三首相の経済政策への期待だ。恩恵が見込まれるとされる人材サービスや建設株の一角が買われている。ただ思惑が先行して資金が集まっている面は否めない。実際に政策効果を取り込む企業はどこか、なお未知数だ。" } ``` Then, in both cases, the field `name` will be detected as Chinese during indexing allowing the search to detect Chinese in queries. Therefore, the query `東京` will be detected as Chinese and only the two last documents will be retrieved by Meilisearch. ## Technical Approach The current PR partially fixes these issues by: 1) Adding a check over potential miss-detections and rerunning the extraction of the document forcing the tokenization over the main Languages detected in it. > 1) run a first extraction allowing the tokenizer to detect any Language in any Script > 2) generate a distribution of tokens by Script and Languages (`script_language`) > 3) if for a Script we have a token distribution of one of the Language that is under the threshold, then we rerun the extraction forbidding the tokenizer to detect the marginal Languages > 4) the tokenizer will fall back on the other available Languages to tokenize the text. For example, if the Chinese were marginally detected compared to the Japanese on the CJ script, then the second extraction will force Japanese tokenization for CJ text in the document. however, the text on another script like Latin will not be impacted by this restriction. 2) Adding a filtering threshold during the search over Languages that have been marginally detected in documents ## Limits This PR introduces 2 arbitrary thresholds: 1) during the indexing, a Language is considered miss-detected if the number of detected tokens of this Language is under 10% of the tokens detected in the same Script (Japanese and Chinese are 2 different Languages sharing the "same" script "CJK"). 2) during the search, a Language is considered marginal if less than 5% of documents are detected as this Language. This PR only partially fixes these issues: - ✅ the query `東京` now find Japanese documents if less than 5% of documents are detected as Chinese. - ✅ the document with the id `105` containing the Japanese field `desc` but the miss-detected field `name` is now completely detected and tokenized as Japanese and is found with the query `東京`. - ❌ the document with the id `4` no longer breaks the search Language detection but continues to be detected as a Chinese document and can't be found during the search. ## Related issue Fixes #3565 ## Possible future enhancements - Change or contribute to the Library used to detect the Language - the related issue on Whatlang: https://github.com/greyblake/whatlang-rs/issues/122 Co-authored-by: curquiza <clementine@meilisearch.com> Co-authored-by: ManyTheFish <many@meilisearch.com> Co-authored-by: Many the fish <many@meilisearch.com>	2023-03-09 15:34:35 +00:00
bors[bot]	48a51e5cd6	Merge #3577 3577: Avoid fetching an LMDB value with an empty string r=ManyTheFish a=Kerollmops # Pull Request ## Related issue Fixes #3574 ## What does this PR do? This PR fixes a bug where an empty key fetches an entry in the database. LMDB throws an error if an empty or too-long key is used to fetch an entry. This empty string seems to have been generated by the Charabia tokenizer. Co-authored-by: Clément Renault <clement@meilisearch.com>	2023-03-09 14:35:25 +00:00
ManyTheFish	2f8eb4f54a	last PR fixes	2023-03-09 15:34:36 +01:00
Many the fish	dea101e3d9	Update meilisearch/src/routes/indexes/mod.rs Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-03-09 15:17:03 +01:00
Clément Renault	175e8a8495	Fix a diacritic issue	2023-03-09 14:57:47 +01:00
Clément Renault	6da54d0cb6	Add a test to fix a diacritic issue	2023-03-09 14:57:38 +01:00
bors[bot]	667bb87e35	Merge #3541 3541: Add cache on the indexes stats r=dureuill a=irevoire Fix https://github.com/meilisearch/meilisearch/issues/3540 Co-authored-by: Tamo <tamo@meilisearch.com> Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-03-09 13:32:52 +00:00
Clément Renault	df48ac8803	Add one more test for the NULL operator	2023-03-09 13:53:37 +01:00
Clément Renault	ff86073288	Add a snapshot for the NULL facet database	2023-03-09 13:32:27 +01:00
Clément Renault	0ad53784e7	Create a new struct to reduce the type complexity	2023-03-09 13:21:21 +01:00
bors[bot]	7935bef4cd	Merge #3567 3567: Clean CI file names r=curquiza a=curquiza Make the CI names more consistent to ease the Gillian's onboarding 😇 No impact for the users or the developers of the team Co-authored-by: curquiza <clementine@meilisearch.com>	2023-03-09 12:20:18 +00:00
Clément Renault	e064c52544	Rename an internal facet deletion method	2023-03-09 13:08:02 +01:00
Clément Renault	e106b16148	Fix a typo in a variable Co-authored-by: Louis Dureuil <louis@meilisearch.com> aaa	2023-03-09 13:08:02 +01:00
Tamo	eddefb0e0f	refactor the error type of the milli::document thing silence a warning	2023-03-09 13:03:14 +01:00
ManyTheFish	dff2715ef3	Try removing needless collect	2023-03-09 11:28:10 +01:00
ManyTheFish	5deea631ea	fix clippy too many arguments	2023-03-09 11:19:13 +01:00
Tamo	c5f22be6e1	add boolean support for csv documents	2023-03-09 11:12:49 +01:00
ManyTheFish	b4b859ec8c	Fix typos	2023-03-09 10:58:35 +01:00
Clément Renault	b1d61f5a02	Add more tests for the NULL filter	2023-03-09 10:04:27 +01:00
curquiza	febc8d1b52	Clean CI file names	2023-03-08 19:12:33 +01:00
Clément Renault	7dc04747fd	Make clippy happy	2023-03-08 17:37:08 +01:00
Clément Renault	7c0cd7172d	Introduce the NULL and NOT value NULL operator	2023-03-08 17:14:34 +01:00
curquiza	b99ef3d336	Update CI to still use ubuntu-18	2023-03-08 17:11:36 +01:00
Clément Renault	43ff236df8	Write the NULL facet values in the database	2023-03-08 16:49:53 +01:00
Clément Renault	19ab4d1a15	Classify the NULL fields values in the facet extractor	2023-03-08 16:49:31 +01:00
Clément Renault	9287858997	Introduce a new facet_id_is_null_docids database in the index	2023-03-08 16:14:00 +01:00
ManyTheFish	7e2fd82e41	Use Language allow list in the highlighter	2023-03-08 12:44:16 +01:00
ManyTheFish	24c0775c67	Change indexing threshold	2023-03-08 12:36:04 +01:00
ManyTheFish	3092cf0448	Fix clippy errors	2023-03-08 10:53:42 +01:00
ManyTheFish	37d4551e8e	Add a threshold filtering the Languages allowed to be detected at search time	2023-03-07 19:38:01 +01:00
ManyTheFish	da48506f15	Rerun extraction when language detection might have failed	2023-03-07 18:35:26 +01:00
Louis Dureuil	2f5b9fbbd8	Restore contribution of the index sizes to the db size - the index size now contributes to the db size even if the index is not authorized	2023-03-07 14:05:27 +01:00
Louis Dureuil	7faa9a22f6	Pass IndexStat by ref in store_stats_of	2023-03-07 14:00:54 +01:00
bors[bot]	370d88f626	Merge #3561 3561: Fix the snapshots permissions on unix system r=irevoire a=irevoire # Pull Request ## Related issue Fixes https://github.com/meilisearch/meilisearch/issues/3507 The snapshot permissions were wrong after the v0.30 and the huge refacto of the index scheduler. Fix this issue + add a test on the permissions on unix Co-authored-by: Tamo <tamo@meilisearch.com>	2023-03-07 08:51:38 +00:00
Tamo	d34faa8f9c	put back the sleep as it was and fix the from	2023-03-06 18:09:09 +01:00
Tamo	e5d0bef6d8	update a comment	2023-03-06 17:04:24 +01:00
Louis Dureuil	76288fad72	Fix snapshots	2023-03-06 16:57:31 +01:00
Louis Dureuil	076a3d371c	Eagerly compute stats as fallback to the cache. - Refactor all around to avoid spawning indexes more times than necessary	2023-03-06 16:57:31 +01:00
Tamo	3bbf760542	update most snapshots	2023-03-06 16:57:31 +01:00
Tamo	fd5c48941a	Add cache on the indexes stats	2023-03-06 16:57:31 +01:00
bors[bot]	df3986cd83	Merge #3510 #3551 #3552 #3553 3510: Add scheduled test to Actions for all features r=curquiza a=jlucktay # Pull Request ## Related issue Fixes #3506. ## What does this PR do? Add a new job to the Rust workflow to run `cargo build` and `cargo test` (on the cron schedule only) with the `--all-features` flag. This will execute across all three environments: Linux, macOS, Windows. Autoformat the Rust workflow file via [the Red Hat YAML extension for Visual Studio Code](https://marketplace.visualstudio.com/items?itemName=redhat.vscode-yaml). This straightens out whitespace and string quoting for safer parsing. As [pointed out by `@irevoire` here](https://github.com/meilisearch/meilisearch/issues/3506#issuecomment-1433501867), changes to CI such as this one will need to wait for #3496 before going ahead. The new action [was executed on my fork](https://github.com/jlucktay/meilisearch/actions/runs/4211694210) but ended up failing on some metrics tests, as called out in that linked comment. ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! 3551: Bump Swatinem/rust-cache from 2.2.0 to 2.2.1 r=curquiza a=dependabot[bot] Bumps [Swatinem/rust-cache](https://github.com/Swatinem/rust-cache) from 2.2.0 to 2.2.1. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/Swatinem/rust-cache/releases">Swatinem/rust-cache's releases</a>.</em></p> <blockquote> <h2>v2.2.1</h2> <ul> <li>Update <code>`@actions/cache</code>` dependency to fix usage of <code>zstd</code> compression.</li> </ul> </blockquote> </details> <details> <summary>Changelog</summary> <p><em>Sourced from <a href="https://github.com/Swatinem/rust-cache/blob/master/CHANGELOG.md">Swatinem/rust-cache's changelog</a>.</em></p> <blockquote> <h2>2.2.1</h2> <ul> <li>Update <code>`@actions/cache</code>` dependency to fix usage of <code>zstd</code> compression.</li> </ul> </blockquote> </details> <details> <summary>Commits</summary> <ul> <li><a href="`6fd3edff69`"><code>6fd3edf</code></a> 2.2.1</li> <li><a href="`a1c019f71a`"><code>a1c019f</code></a> update dependencies and rebuild</li> <li><a href="`664ce0090f`"><code>664ce00</code></a> chore: Create check-dist.yml (<a href="https://github-redirect.dependabot.com/Swatinem/rust-cache/issues/96">#96</a>)</li> <li>See full diff in <a href="https://github.com/Swatinem/rust-cache/compare/v2.2.0...v2.2.1">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=Swatinem/rust-cache&package-manager=github_actions&previous-version=2.2.0&new-version=2.2.1)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) You can trigger a rebase of this PR by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) </details> 3552: Bump svenstaro/upload-release-action from 2.4.0 to 2.5.0 r=curquiza a=dependabot[bot] Bumps [svenstaro/upload-release-action](https://github.com/svenstaro/upload-release-action) from 2.4.0 to 2.5.0. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/svenstaro/upload-release-action/releases">svenstaro/upload-release-action's releases</a>.</em></p> <blockquote> <h2>2.5.0</h2> <ul> <li>Add retry to upload release <a href="https://github-redirect.dependabot.com/svenstaro/upload-release-action/pull/96">#96</a> (thanks <a href="https://github.com/sonphantrung"><code>`@sonphantrung</code></a>)</li>` </ul> <h2>2.4.1</h2> <ul> <li>Modernize octokit usage</li> </ul> </blockquote> </details> <details> <summary>Changelog</summary> <p><em>Sourced from <a href="https://github.com/svenstaro/upload-release-action/blob/master/CHANGELOG.md">svenstaro/upload-release-action's changelog</a>.</em></p> <blockquote> <h2>[2.5.0] - 2023-02-21</h2> <ul> <li>Add retry to upload release <a href="https://github-redirect.dependabot.com/svenstaro/upload-release-action/pull/96">#96</a> (thanks <a href="https://github.com/sonphantrung"><code>`@sonphantrung</code></a>)</li>` </ul> <h2>[2.4.1] - 2023-02-01</h2> <ul> <li>Modernize octokit usage</li> </ul> </blockquote> </details> <details> <summary>Commits</summary> <ul> <li><a href="`7319e4733e`"><code>7319e47</code></a> 2.5.0</li> <li><a href="`4e86b8565b`"><code>4e86b85</code></a> Prepare release</li> <li><a href="`3a6baf0f12`"><code>3a6baf0</code></a> Add CHANGELOG entry for retry feature</li> <li><a href="`e8c797e08e`"><code>e8c797e</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/svenstaro/upload-release-action/issues/96">#96</a> from sonphantrung/retry-v2</li> <li><a href="`cf83be2c7f`"><code>cf83be2</code></a> Merge branch 'master' into retry-v2</li> <li><a href="`cfdd9b50bd`"><code>cfdd9b5</code></a> Merge branch 'retry' of <a href="https://github.com/messense/upload-release-action">https://github.com/messense/upload-release-action</a></li> <li><a href="`cc92c9093e`"><code>cc92c90</code></a> 2.4.1</li> <li><a href="`72f6bf584a`"><code>72f6bf5</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/svenstaro/upload-release-action/issues/93">#93</a> from ggreif/gabor/fix</li> <li><a href="`f2899b4677`"><code>f2899b4</code></a> use <code>createReadStream</code></li> <li><a href="`af306bddfe`"><code>af306bd</code></a> Revert "use the <code>`@file</code>` mechanism of octokit-5"</li> <li>Additional commits viewable in <a href="https://github.com/svenstaro/upload-release-action/compare/2.4.0...2.5.0">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=svenstaro/upload-release-action&package-manager=github_actions&previous-version=2.4.0&new-version=2.5.0)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) You can trigger a rebase of this PR by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) </details> 3553: Bump docker/build-push-action from 3 to 4 r=curquiza a=dependabot[bot] Bumps [docker/build-push-action](https://github.com/docker/build-push-action) from 3 to 4. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/docker/build-push-action/releases">docker/build-push-action's releases</a>.</em></p> <blockquote> <h2>v4.0.0</h2> <blockquote> <p><strong>Note</strong></p> <p>Buildx v0.10 enables support for a minimal <a href="https://slsa.dev/provenance/">SLSA Provenance</a> attestation, which requires support for <a href="https://github.com/opencontainers/image-spec">OCI-compliant</a> multi-platform images. This may introduce issues with registry and runtime support (e.g. <a href="https://github-redirect.dependabot.com/docker/buildx/issues/1533">Google Cloud Run and AWS Lambda</a>). You can optionally disable the default provenance attestation functionality using <code>provenance: false</code>.</p> </blockquote> <ul> <li>Revert disable provenance by default if not set by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` in <a href="https://github-redirect.dependabot.com/docker/build-push-action/pull/784">docker/build-push-action#784</a></li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/build-push-action/compare/v3.3.1...v4.0.0">https://github.com/docker/build-push-action/compare/v3.3.1...v4.0.0</a></p> <h2>v3.3.1</h2> <ul> <li>Disable provenance by default if not set by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/781">#781</a>)</li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/build-push-action/compare/v3.3.0...v3.3.1">https://github.com/docker/build-push-action/compare/v3.3.0...v3.3.1</a></p> <h2>v3.3.0</h2> <blockquote> <p><strong>Note</strong></p> <p>Buildx v0.10 enables support for a minimal <a href="https://slsa.dev/provenance/">SLSA Provenance</a> attestation, which requires support for <a href="https://github.com/opencontainers/image-spec">OCI-compliant</a> multi-platform images. This may introduce issues with registry and runtime support (e.g. <a href="https://github-redirect.dependabot.com/docker/buildx/issues/1533">Google Cloud Run and AWS Lambda</a>). You can optionally disable the default provenance attestation functionality using <code>provenance: false</code>.</p> </blockquote> <ul> <li>Add <code>attests</code>, <code>provenance</code> and <code>sbom</code> inputs by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/746">#746</a> <a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/759">#759</a>)</li> <li>Log GitHub Actions runtime token access controls by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/707">#707</a>)</li> <li>Examples moved to <a href="https://docs.docker.com/build/ci/github-actions/examples/">docs website</a> by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/718">#718</a>)</li> <li>Bump minimatch from 3.0.4 to 3.1.2 (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/732">#732</a>)</li> <li>Bump csv-parse from 5.3.0 to 5.3.3 (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/729">#729</a>)</li> <li>Bump json5 from 2.2.0 to 2.2.3 (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/749">#749</a>)</li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/build-push-action/compare/v3.2.0...v3.3.0">https://github.com/docker/build-push-action/compare/v3.2.0...v3.3.0</a></p> <h2>v3.2.0</h2> <ul> <li>Remove workaround for <code>setOutput</code> by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/704">#704</a>)</li> <li>Docs: fix Git context link and add more details about subdir support by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/685">#685</a>)</li> <li>Docs: named context by <a href="https://github.com/baibaratsky"><code>`@baibaratsky</code></a>` and <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/665">#665</a>)</li> <li>Bump <code>`@actions/core</code>` from 1.9.0 to 1.10.0 (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/667">#667</a> <a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/695">#695</a>)</li> <li>Bump <code>`@actions/github</code>` from 5.0.3 to 5.1.1 (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/696">#696</a>)</li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/build-push-action/compare/v3.1.1...v3.2.0">https://github.com/docker/build-push-action/compare/v3.1.1...v3.2.0</a></p> <h2>v3.1.1</h2> <ul> <li>Fix GitHub token not passed with Git context if subdir defined by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/663">#663</a>)</li> <li>Replace deprecated <code>fs.rmdir</code> with <code>fs.rm</code> by <a href="https://github.com/bendrucker"><code>`@bendrucker</code></a>` (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/657">#657</a>)</li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/build-push-action/compare/v3.1.0...v3.1.1">https://github.com/docker/build-push-action/compare/v3.1.0...v3.1.1</a></p> <h2>v3.1.0</h2> <ul> <li><code>no-cache-filters</code> input by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/653">#653</a>)</li> <li>Bump <code>`@actions/github</code>` from 5.0.1 to 5.0.3 (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/619">#619</a>)</li> <li>Bump <code>`@actions/core</code>` from 1.6.0 to 1.9.0 (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/620">#620</a> <a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/637">#637</a>)</li> <li>Bump csv-parse from 5.0.4 to 5.3.0 (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/623">#623</a> <a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/650">#650</a>)</li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/build-push-action/compare/v3.0.0...v3.1.0">https://github.com/docker/build-push-action/compare/v3.0.0...v3.1.0</a></p> </blockquote> </details> <details> <summary>Commits</summary> <ul> <li><a href="`3b5e8027fc`"><code>3b5e802</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/784">#784</a> from crazy-max/enable-provenance</li> <li><a href="`02d3266a89`"><code>02d3266</code></a> update generated content</li> <li><a href="`f403dafe18`"><code>f403daf</code></a> revert disable provenance by default if not set</li> <li>See full diff in <a href="https://github.com/docker/build-push-action/compare/v3...v4">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=docker/build-push-action&package-manager=github_actions&previous-version=3&new-version=4)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) You can trigger a rebase of this PR by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) </details> Co-authored-by: James Lucktaylor <jlucktay+github@gmail.com> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2023-03-06 15:45:33 +00:00
Tamo	e704728ee7	fix the snapshots permissions on unix system	2023-03-06 16:28:40 +01:00
bors[bot]	34ed6518ae	Merge #3554 3554: Bump docker/metadata-action from 3 to 4 r=curquiza a=dependabot[bot] Bumps [docker/metadata-action](https://github.com/docker/metadata-action) from 3 to 4. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/docker/metadata-action/releases">docker/metadata-action's releases</a>.</em></p> <blockquote> <h2>v4.0.0</h2> <ul> <li>Node 16 as default runtime by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/176">#176</a>) <ul> <li>This requires a minimum <a href="https://github.com/actions/runner/releases/tag/v2.285.0">Actions Runner</a> version of v2.285.0, which is by default available in GHES 3.4 or later.</li> </ul> </li> <li>Do not sanitize before pattern matching by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/201">#201</a>) <ul> <li>Breaking change with <code>type=match</code> pattern matching</li> </ul> </li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/metadata-action/compare/v3.8.0...v4.0.0">https://github.com/docker/metadata-action/compare/v3.8.0...v4.0.0</a></p> <h2>v3.8.0</h2> <ul> <li>Add attribute to enable/disable images by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/193">#193</a>)</li> <li>Add <code>is_default_branch</code> global expression by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/192">#192</a> <a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/197">#197</a> <a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/198">#198</a>)</li> <li>Update fixtures (dev) by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/190">#190</a>)</li> <li>Bump semver from 7.3.5 to 7.3.7 (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/185">#185</a>)</li> <li>Bump moment from 2.29.2 to 2.29.3 (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/187">#187</a>)</li> <li>Bump csv-parse from 4.16.3 to 5.0.4 (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/195">#195</a>)</li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/metadata-action/compare/v3.7.0...v3.8.0">https://github.com/docker/metadata-action/compare/v3.7.0...v3.8.0</a></p> <h2>v3.7.0</h2> <ul> <li>Handle comments for multi-line inputs (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/172">#172</a>)</li> <li>Missing <code>json</code> output in <code>action.yml</code> (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/167">#167</a>)</li> <li>Update dev dependencies and workflow (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/175">#175</a>)</li> <li>Bump minimist from 1.2.5 to 1.2.6 (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/182">#182</a>)</li> <li>Bump moment from 2.29.1 to 2.29.2 (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/180">#180</a>)</li> <li>Bump <code>`@actions/github</code>` from 5.0.0 to 5.0.1 (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/179">#179</a>)</li> <li>Bump node-fetch from 2.6.1 to 2.6.7 (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/173">#173</a>)</li> </ul> <h2>v3.6.2</h2> <ul> <li>Handle raw statement for pre-release (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/155">#155</a> <a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/156">#156</a>)</li> </ul> <h2>v3.6.1</h2> <ul> <li>Preserve quotes inside unquoted field (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/153">#153</a>)</li> </ul> <h2>v3.6.0</h2> <ul> <li><code>base_ref</code> global expression (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/142">#142</a>)</li> <li>Trim tags and flavor inputs (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/143">#143</a>)</li> <li>Bump <code>`@actions/core</code>` from 1.5.0 to 1.6.0 (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/135">#135</a>)</li> <li>Bump ansi-regex from 5.0.0 to 5.0.1 (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/134">#134</a>)</li> <li>Bump tmpl from 1.0.4 to 1.0.5 (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/132">#132</a>)</li> <li>Bump csv-parse from 4.16.0 to 4.16.3 (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/131">#131</a>)</li> </ul> <h2>v3.5.0</h2> <ul> <li>Add global expression <code>date</code> (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/121">#121</a>)</li> <li>Bump <code>`@actions/core</code>` from 1.4.0 to 1.5.0 (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/122">#122</a>)</li> </ul> <h2>v3.4.1</h2> <ul> <li>Only return edge if branch matches (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/115">#115</a>)</li> </ul> <h2>v3.4.0</h2> <ul> <li>PEP 440 support (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/108">#108</a>)</li> </ul> <!-- raw HTML omitted --> </blockquote> <p>... (truncated)</p> </details> <details> <summary>Upgrade guide</summary> <p><em>Sourced from <a href="https://github.com/docker/metadata-action/blob/master/UPGRADE.md">docker/metadata-action's upgrade guide</a>.</em></p> <blockquote> <h1>Upgrade notes</h1> <h2>v2 to v3</h2> <ul> <li>Repository has been moved to docker org. Replace <code>crazy-max/ghaction-docker-meta@v2</code> with <code>docker/metadata-action@v4</code></li> <li>The default bake target has been changed: <code>ghaction-docker-meta</code> > <code>docker-metadata-action</code></li> </ul> <h2>v1 to v2</h2> <ul> <li><a href="https://github.com/docker/metadata-action/blob/master/#inputs">inputs</a> <ul> <li><a href="https://github.com/docker/metadata-action/blob/master/#tag-sha"><code>tag-sha</code></a></li> <li><a href="https://github.com/docker/metadata-action/blob/master/#tag-edge--tag-edge-branch"><code>tag-edge</code> / <code>tag-edge-branch</code></a></li> <li><a href="https://github.com/docker/metadata-action/blob/master/#tag-semver"><code>tag-semver</code></a></li> <li><a href="https://github.com/docker/metadata-action/blob/master/#tag-match--tag-match-group"><code>tag-match</code> / <code>tag-match-group</code></a></li> <li><a href="https://github.com/docker/metadata-action/blob/master/#tag-latest"><code>tag-latest</code></a></li> <li><a href="https://github.com/docker/metadata-action/blob/master/#tag-schedule"><code>tag-schedule</code></a></li> <li><a href="https://github.com/docker/metadata-action/blob/master/#tag-custom--tag-custom-only"><code>tag-custom</code> / <code>tag-custom-only</code></a></li> <li><a href="https://github.com/docker/metadata-action/blob/master/#label-custom"><code>label-custom</code></a></li> </ul> </li> <li><a href="https://github.com/docker/metadata-action/blob/master/#basic-workflow">Basic workflow</a></li> <li><a href="https://github.com/docker/metadata-action/blob/master/#semver-workflow">Semver workflow</a></li> </ul> <h3>inputs</h3> <table> <thead> <tr> <th>New</th> <th>Unchanged</th> <th>Removed</th> </tr> </thead> <tbody> <tr> <td><code>tags</code></td> <td><code>images</code></td> <td><code>tag-sha</code></td> </tr> <tr> <td><code>flavor</code></td> <td><code>sep-tags</code></td> <td><code>tag-edge</code></td> </tr> <tr> <td><code>labels</code></td> <td><code>sep-labels</code></td> <td><code>tag-edge-branch</code></td> </tr> <tr> <td></td> <td></td> <td><code>tag-semver</code></td> </tr> <tr> <td></td> <td></td> <td><code>tag-match</code></td> </tr> <tr> <td></td> <td></td> <td><code>tag-match-group</code></td> </tr> <tr> <td></td> <td></td> <td><code>tag-latest</code></td> </tr> <tr> <td></td> <td></td> <td><code>tag-schedule</code></td> </tr> <tr> <td></td> <td></td> <td><code>tag-custom</code></td> </tr> <tr> <td></td> <td></td> <td><code>tag-custom-only</code></td> </tr> <tr> <td></td> <td></td> <td><code>label-custom</code></td> </tr> </tbody> </table> <h4><code>tag-sha</code></h4> <pre lang="yaml"><code>tags: \| type=sha </code></pre> <h4><code>tag-edge</code> / <code>tag-edge-branch</code></h4> <pre lang="yaml"><code>tags: \| # default branch </tr></table> </code></pre> </blockquote> <p>... (truncated)</p> </details> <details> <summary>Commits</summary> <ul> <li><a href="`507c2f2dc5`"><code>507c2f2</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/257">#257</a> from crazy-max/env-output</li> <li><a href="`04861f5102`"><code>04861f5</code></a> update generated content</li> <li><a href="`6729545cde`"><code>6729545</code></a> Provide outputs as env vars</li> <li><a href="`05d22bf317`"><code>05d22bf</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/256">#256</a> from crazy-max/fix-readme</li> <li><a href="`70b403b46b`"><code>70b403b</code></a> Fix README</li> <li><a href="`9e6ae02878`"><code>9e6ae02</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/252">#252</a> from docker/dependabot/npm_and_yarn/json5-2.2.3</li> <li><a href="`3d239e8b8a`"><code>3d239e8</code></a> Bump json5 from 2.2.0 to 2.2.3</li> <li><a href="`7cb52e2750`"><code>7cb52e2</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/251">#251</a> from chroju/set_timezone</li> <li><a href="`90a1d5cf21`"><code>90a1d5c</code></a> Add tz attribute to handlebar date function</li> <li><a href="`c98ac5e987`"><code>c98ac5e</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/249">#249</a> from crazy-max/fix-readme</li> <li>Additional commits viewable in <a href="https://github.com/docker/metadata-action/compare/v3...v4">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=docker/metadata-action&package-manager=github_actions&previous-version=3&new-version=4)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) You can trigger a rebase of this PR by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) </details> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2023-03-06 14:56:20 +00:00
bors[bot]	c0ede6d152	Merge #3562 3562: Update version for the next release (v1.1.0) in Cargo.toml r=curquiza a=meili-bot ⚠️ This PR is automatically generated. Check the new version is the expected one and Cargo.lock has been updated before merging. Co-authored-by: curquiza <curquiza@users.noreply.github.com>	2023-03-06 13:54:16 +00:00
curquiza	577e7126f9	Update version for the next release (v1.1.0) in Cargo.toml	2023-03-06 13:52:54 +00:00
James Lucktaylor	22219fd88f	ci(actions/rust): explicitly set up dependencies and toolchain override	2023-03-06 12:45:08 +00:00
James Lucktaylor	a9e17ab8c6	style(actions/rust): resolve PR review	2023-03-03 12:08:30 +00:00
James Lucktaylor	2dd948a4a1	ci(actions/rust): align with test-linux job	2023-03-03 12:07:42 +00:00
James Lucktaylor	76cf1bff87	Add scheduled test to Actions for all features Add a new job to the Rust workflow to run 'cargo build' and 'cargo test' (on the cron schedule only) with the '--all-features' flag. This will execute across all three environments: Linux, macOS, Windows. Autoformat the Rust workflow file via the Red Hat YAML extension for Visual Studio Code: https://marketplace.visualstudio.com/items?itemName=redhat.vscode-yaml This straightens out whitespace and string quoting for safer parsing. Fixes #3506.	2023-03-03 12:01:14 +00:00
bors[bot]	3d1046369c	Merge #3529 3529: Add an analytics on the geo bounding box feature r=ManyTheFish a=irevoire Fixes #3527 [The specification of the geoBoundingBox](https://github.com/meilisearch/specifications/pull/223) feature has been updated and now introduces a new analytics to follow the usage of the geoBoundingBox feature in the search requests. Co-authored-by: Tamo <tamo@meilisearch.com>	2023-03-02 11:58:39 +00:00
bors[bot]	4f1ccbc495	Merge #3525 3525: Fix phrase search containing stop words r=ManyTheFish a=ManyTheFish # Summary A search with a phrase containing only stop words was returning an HTTP error 500, this PR filters the phrase containing only stop words dropping them before the search starts, a query with a phrase containing only stop words now behaves like a placeholder search. fixes https://github.com/meilisearch/meilisearch/issues/3521 related v1.0.2 PR on milli: https://github.com/meilisearch/milli/pull/779 Co-authored-by: ManyTheFish <many@meilisearch.com>	2023-03-02 10:55:37 +00:00
ManyTheFish	37489fd495	Return an internal error in the case of matching word is invalid	2023-03-01 19:05:16 +01:00
dependabot[bot]	c0d8eb295d	Bump docker/metadata-action from 3 to 4 Bumps [docker/metadata-action](https://github.com/docker/metadata-action) from 3 to 4. - [Release notes](https://github.com/docker/metadata-action/releases) - [Upgrade guide](https://github.com/docker/metadata-action/blob/master/UPGRADE.md) - [Commits](https://github.com/docker/metadata-action/compare/v3...v4) --- updated-dependencies: - dependency-name: docker/metadata-action dependency-type: direct:production update-type: version-update:semver-major ... Signed-off-by: dependabot[bot] <support@github.com>	2023-03-01 17:58:18 +00:00
dependabot[bot]	bcd3f6054a	Bump docker/build-push-action from 3 to 4 Bumps [docker/build-push-action](https://github.com/docker/build-push-action) from 3 to 4. - [Release notes](https://github.com/docker/build-push-action/releases) - [Commits](https://github.com/docker/build-push-action/compare/v3...v4) --- updated-dependencies: - dependency-name: docker/build-push-action dependency-type: direct:production update-type: version-update:semver-major ... Signed-off-by: dependabot[bot] <support@github.com>	2023-03-01 17:58:11 +00:00
dependabot[bot]	3a0314f9de	Bump svenstaro/upload-release-action from 2.4.0 to 2.5.0 Bumps [svenstaro/upload-release-action](https://github.com/svenstaro/upload-release-action) from 2.4.0 to 2.5.0. - [Release notes](https://github.com/svenstaro/upload-release-action/releases) - [Changelog](https://github.com/svenstaro/upload-release-action/blob/master/CHANGELOG.md) - [Commits](https://github.com/svenstaro/upload-release-action/compare/2.4.0...2.5.0) --- updated-dependencies: - dependency-name: svenstaro/upload-release-action dependency-type: direct:production update-type: version-update:semver-minor ... Signed-off-by: dependabot[bot] <support@github.com>	2023-03-01 17:58:05 +00:00
dependabot[bot]	fa4d8b8348	Bump Swatinem/rust-cache from 2.2.0 to 2.2.1 Bumps [Swatinem/rust-cache](https://github.com/Swatinem/rust-cache) from 2.2.0 to 2.2.1. - [Release notes](https://github.com/Swatinem/rust-cache/releases) - [Changelog](https://github.com/Swatinem/rust-cache/blob/master/CHANGELOG.md) - [Commits](https://github.com/Swatinem/rust-cache/compare/v2.2.0...v2.2.1) --- updated-dependencies: - dependency-name: Swatinem/rust-cache dependency-type: direct:production update-type: version-update:semver-patch ... Signed-off-by: dependabot[bot] <support@github.com>	2023-03-01 17:57:57 +00:00
bors[bot]	d9e19c89c5	Merge #3544 3544: Attempt to use default vram budget for faster startup r=Kerollmops a=dureuill # Pull Request ## Related issue Follow-up to #3382: addresses the added startup time on Windows/macOS. ## What does this PR do? - Attempt to skip budget calculation by using "known good values" instead - Perform dichotomic budget calculation as fallback only when the known value is not actually good. Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-03-01 09:49:38 +00:00
bors[bot]	18bf740ee2	Merge #3539 3539: Update migration link to the docs r=curquiza a=curquiza Fixes https://github.com/meilisearch/meilisearch/issues/3449 Co-authored-by: curquiza <clementine@meilisearch.com>	2023-02-28 18:21:11 +00:00
Louis Dureuil	0202ff8ab4	Attempt to use default budget for faster startup	2023-02-28 10:55:43 +01:00
bors[bot]	fbe4ab158e	Merge #3543 3543: config: case `experimental_enable_metrics` in snake_case r=dureuill a=dureuill # Pull Request Avoids "Error: unknown field `experimental-enable-metrics` at line 1 column 1" error when using the default config file. Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-02-28 09:35:25 +00:00
Louis Dureuil	92318ca573	config: case `experimental_enable_metrics` in snake_case	2023-02-27 17:14:06 +01:00
bors[bot]	6ca7a109b9	Merge #3538 3538: Improve the api key of the metrics r=dureuill a=irevoire Related to https://github.com/meilisearch/meilisearch/pull/3524#discussion_r1115903998 Update: https://github.com/meilisearch/meilisearch/issues/3523 Right after merging the PR, we changed our minds and decided to update the way we handle the API keys on the metrics route. Now instead of bypassing all the applied rules of the API key, we forbid the usage of the `/metrics` route if you have any restrictions on the indexes. Co-authored-by: Tamo <tamo@meilisearch.com> Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-02-27 13:46:57 +00:00
Louis Dureuil	d4d4702f1b	Rephrase hint message	2023-02-27 13:46:16 +01:00
curquiza	2648bbca25	Update migration link to the docs	2023-02-23 18:36:30 +01:00
bors[bot]	562c86ea01	Merge #3519 3519: Update comments in version bump CI r=irevoire a=curquiza Following https://github.com/meilisearch/meilisearch/pull/3499 Co-authored-by: curquiza <clementine@meilisearch.com>	2023-02-23 16:46:10 +00:00
Tamo	7ae10abb6b	fix the auth tests	2023-02-23 17:27:42 +01:00
Tamo	dc533584c6	Forbid the usage of the metrics route if your API key have a limitation on the indexes	2023-02-23 17:13:22 +01:00
bors[bot]	442c1e36de	Merge #3537 3537: fix a bug where the filestore could try to parse its own tmp file and fail (main) r=irevoire a=curquiza Co-authored-by: Tamo <tamo@meilisearch.com>	2023-02-23 15:58:05 +00:00
Tamo	66b5e4b548	fix a bug where the filestore could try to parse its own tmp file and fail	2023-02-23 16:52:41 +01:00
bors[bot]	89ac1015f3	Merge #3524 3524: Update the metrics route r=irevoire a=irevoire Fixes #3523 Make the metrics available by default without a feature flag. + Rename the cli-flag to `experimental-enable-metrics`. Co-authored-by: Tamo <tamo@meilisearch.com>	2023-02-23 15:11:10 +00:00
bors[bot]	ca25904c26	Merge #3331 3331: Limit the number of concurrently opened indexes r=dureuill a=dureuill # Pull Request ## Related issue Relevant to #1841, fixes #3382 ## What does this PR do? ### User standpoint - Limit the number of concurrently opened indexes (currently, the number of indexes that can be concurrently opened is computed at startup) - When too many an index is opened, the least recently used one is closed and its virtual memory released. - This allows a user to have an arbitrary number of indexes of an arbitrary size ### Implementation standpoint - Added a LRU cache map in `index-scheduler::lru`. A more complete implementation (eg with helper functions not used here) is available but would better fit a dedicated crate. - Use the LRU cache map in the `IndexScheduler`. To simplify the lifecycle of indexes, they are never removed from the cache when they are in the middle of a resize or delete operation. To achieve this, an intermediate `Vec` stores the UUIDs of the indexes that are in the middle of such an operation. - Upon creating the index scheduler object, compute the total virtual memory that is adressable by using a dichotomic search on the max size of an index. Use this as a base to compute the number of indexes that can be open with 2TiB per index. If the virtual memory address space is lower than 2TiB, then only allow for 1 index of a fraction of that size. Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-02-23 14:20:52 +00:00
Tamo	8a1b1a95f3	comment the right of the metrics	2023-02-23 13:59:01 +01:00
Tamo	8d47d2d018	update the auth api after the rebase	2023-02-23 13:15:51 +01:00
Tamo	5082cd5e67	update the config file to mention the experimental metrics feature	2023-02-23 12:26:22 +01:00
Tamo	750a2b6842	Update meilisearch/src/option.rs Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-02-23 12:26:22 +01:00
Tamo	bc7d4112d9	send the cli experimental feature in the analytics	2023-02-23 12:26:22 +01:00
Tamo	88a18677d0	rename the metrics cli flag	2023-02-23 12:26:22 +01:00
Tamo	68e30214ca	remove the feature flag and reorganize the module slightly	2023-02-23 12:26:21 +01:00
bors[bot]	b985b96e4e	Merge #3530 3530: Fix highlighter bug r=Kerollmops a=ManyTheFish # Pull Request There was a highlighting issue on CJK's character, we were highlighting too many characters and these additional characters were duplicated after the highlight tag. ## Related issue Fixes #3517 Fixes #3526 ## What does this PR do? - add a test showcasing the bug - fix the bug by activating the char_map creation of the tokenizer during the highlighting process Co-authored-by: ManyTheFish <many@meilisearch.com>	2023-02-23 10:59:43 +00:00
Louis Dureuil	71e7900c67	move index_map to file	2023-02-23 11:29:11 +01:00
Louis Dureuil	431782f3ee	Move index_mapper to mod.rs	2023-02-23 11:29:11 +01:00
Louis Dureuil	3db613ff77	Don't iterate all indexes manually	2023-02-23 11:29:09 +01:00
Louis Dureuil	5822764be9	Skip computing index budget in tests	2023-02-23 11:23:39 +01:00
Louis Dureuil	c63294f331	Switch to 2TiB default index size, updates documentation	2023-02-23 11:23:39 +01:00
Louis Dureuil	a529bf160c	Compute budget	2023-02-23 11:23:39 +01:00
Louis Dureuil	f1119f2dc2	Add dichotomic search to utils	2023-02-23 11:23:39 +01:00
Louis Dureuil	1db7d5d851	Add basic tests for index eviction and resize	2023-02-23 11:23:39 +01:00
Louis Dureuil	80b060f920	Use LRU cache	2023-02-23 11:23:39 +01:00
Louis Dureuil	fdf043580c	Add LruMap	2023-02-23 11:23:38 +01:00
bors[bot]	f62703cd67	Merge #3534 3534: Update the csv error code from InvalidIndexCsvDelimiter to InvalidDocumentCsvDelimiter r=Kerollmops a=irevoire Fixes #3533 Co-authored-by: Tamo <tamo@meilisearch.com>	2023-02-23 07:05:12 +00:00
Tamo	76f82c880d	update the csv error code from InvalidIndexCsvDelimiter to InvalidDocumentCsvDelimiter	2023-02-22 19:26:48 +01:00
bors[bot]	6eeba3a8ab	Merge #3417 3417: Allow multiple searches in a single request r=irevoire a=dureuill # Pull Request ## Related issue Fixes #3427 ## What does this PR do? ### User standpoint - Adds a new `/multi-search` entry point (not to be confused with the existing `/{index_uid}/search` entry points) that accepts a POST whose body is an object containing an array of queries. - Each query must specify on which index it acts by providing its `indexUid`. Other parameters are identical to the one in the existing search routes (`q`, `limit`, etc.). - The response is a JSON object containing an array of the results for each search query as if it had been performed using the `/{index_uid}/search` routes. ### Implementation standpoint - Refactor authentication module: - Allow tenant token to be checked even without an index in URL - Add `meilisearch-auth` as a dependency to `index-scheduler` so as to have a working method of checking if the indexes are authorized there that takes into account both the API key and the tenant token (existing method relied on a behavior that was returning the allowed indexes from the API key as long as there weren't any tenant token) - Make `AuthFilter` an object with invariants and so its fields are now private - Use the methods of `AuthFilter` to know if an index is authorized rather than relying on its internal search rules. - Make tenant token search rules optional and `None` when the `AuthFilter` was not built with a tenant token. - Add a new `routes::index::search::multiple_search` module containing a post handler that performs the same work as the existing `routes::index::search` post handler, but in a loop. - Add various tests - Add authentication test suite ### Sample request <details> <summary> Click to see request/response </summary> ```json ~/datasets ❯ curl \ -X POST 'http://localhost:7700/multi-search' \ -H 'Content-Type: application/json' \ --data-binary '{"queries": [{ "indexUid": "index-0", "q": "toto", "limit": 1 }, {"indexUid": "index-1", "q": "titi", "limit": 1}]}' \| jsonxf { "results": [ { "indexUid": "index-0", "hits": [ { "id": 20480, "title": "Toto - 25th Anniversary - Live in Amsterdam", "overview": "Filmed in High Definition in Amsterdam on Toto's 25th Anniversary Tour in 2003, this stunning concert captures the band at their very best, reunited with original vocalist Bobby Kimball. The set combines all their hits with tracks from their latest album \"Through the Looking Glass\" and other live favorites, performed in front of a wildly enthusastic sell-out crowd. Extras include 35 minute behind-the-scenes film following the band through various stages of their world tour including footage from Japan, Thailand, South Korea, and France. Toto celebrate their 25th anniversary with this blistering live concert, filmed in Amsterdam on February 25th, 2003. Proving they've still got exactly what it takes to move a crowd, the band perform a mixture of medley's, solo spots, and huge hits. Tracks include \"Rosanna,\" \"Africa,\" \"Hold The Line,\" a cover of the Beatles' \"While My Guitar Gently Weeps,\" and many more.", "genres": [ "Music" ], "poster": "https://image.tmdb.org/t/p/w500/7SCbUPwoB8Z7VUIA1Rn1WWwjNiT.jpg", "release_date": 1064275200 } ], "query": "toto", "processingTimeMs": 1, "limit": 1, "offset": 0, "estimatedTotalHits": 17 }, { "indexUid": "index-1", "hits": [ { "id": 41212, "title": "Titicut Follies", "overview": "The film is a stark and graphic portrayal of the conditions that existed at the State Prison for the Criminally Insane at Bridgewater, Massachusetts. TITICUT FOLLIES documents the various ways the inmates are treated by the guards, social workers and psychiatrists.", "genres": [ "Documentary" ], "poster": "https://image.tmdb.org/t/p/w500/2Ju5hn1ofOPeP1eRJtQWakiHuhW.jpg", "release_date": -70934400 } ], "query": "titi", "processingTimeMs": 0, "limit": 1, "offset": 0, "estimatedTotalHits": 7 } ]} ``` </details> ## PR checklist Please check if your PR fulfills the following requirements: - [ ] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [ ] Have you read the contributing guidelines? - [ ] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-02-22 17:23:45 +00:00
ManyTheFish	28d6a4466d	Make the tokenizer creating a char map during highlighting	2023-02-22 17:43:10 +01:00
Louis Dureuil	1ba2fae3ae	multi-search/authentication: Add authentication tests	2023-02-22 17:04:12 +01:00
Louis Dureuil	28d6ab78de	multi-search: Add multi search tests	2023-02-22 17:04:12 +01:00
Louis Dureuil	3ba5dfb6ec	multi-search: Add test server search method for multi search	2023-02-22 17:04:12 +01:00
Louis Dureuil	a23fbf6c7b	multi-search: Add search with an array of indexes	2023-02-22 17:04:12 +01:00
Louis Dureuil	596a98f7c6	multi-search: Add basic analytics	2023-02-22 16:37:18 +01:00
Louis Dureuil	14c4a222da	Authentication: AuthFilter::allow_index_creation both check that the index is authorized and the IndexCreate action	2023-02-22 16:37:13 +01:00
Louis Dureuil	690bb2e5cc	Authentication: Make allow_index_creation a private field	2023-02-22 16:35:52 +01:00
Louis Dureuil	d0f2c9c72e	Authentication: Make search_rules optional in AuthFilter	2023-02-22 16:35:52 +01:00
Louis Dureuil	42577403d8	Authentication: Directly pass the authfilter to the index scheduler	2023-02-22 16:35:52 +01:00
Louis Dureuil	c8c5944094	Authentication: is_index_authorized takes into account API key indexes even with a tenant token	2023-02-22 16:35:52 +01:00
Louis Dureuil	4b65851793	Authentication: Refactor authentication check to work for tenant token even without an index in URL Callers need to manually check `is_index_authorized` when using the route without an index in URL	2023-02-22 16:35:51 +01:00
Louis Dureuil	10d4a1a9af	Make ResponseError code and message pub so that they can be modified	2023-02-22 16:35:51 +01:00
ManyTheFish	ad35edfa32	Add test	2023-02-22 15:47:15 +01:00
Tamo	033417e9cc	add an analytics on the geo bounding box feature	2023-02-22 15:35:26 +01:00
bors[bot]	ac5a1e4c4b	Merge #3423 3423: Add min and max facet stats r=dureuill a=dureuill # Pull Request ## Related issue Fixes #3426 ## What does this PR do? ### User standpoint - When using a `facets` parameter in search, the facets that have numeric values are displayed in a new section of the response called `facetStats` that contains, per facet, the numeric min and max value of the hits returned by the search. <details> <summary> Sample request/response </summary> ```json ❯ curl \ -X POST 'http://localhost:7700/indexes/meteorites/search?facets=mass' \ -H 'Content-Type: application/json' \ --data-binary '{ "q": "LL6", "facets":["mass", "recclass"], "limit": 5 }' \| jsonxf { "hits": [ { "name": "Niger (LL6)", "id": "16975", "nametype": "Valid", "recclass": "LL6", "mass": 3.3, "fall": "Fell" }, { "name": "Appley Bridge", "id": "2318", "nametype": "Valid", "recclass": "LL6", "mass": 15000, "fall": "Fell", "_geo": { "lat": 53.58333, "lng": -2.71667 } }, { "name": "Athens", "id": "4885", "nametype": "Valid", "recclass": "LL6", "mass": 265, "fall": "Fell", "_geo": { "lat": 34.75, "lng": -87.0 } }, { "name": "Bandong", "id": "4935", "nametype": "Valid", "recclass": "LL6", "mass": 11500, "fall": "Fell", "_geo": { "lat": -6.91667, "lng": 107.6 } }, { "name": "Benguerir", "id": "30443", "nametype": "Valid", "recclass": "LL6", "mass": 25000, "fall": "Fell", "_geo": { "lat": 32.25, "lng": -8.15 } } ], "query": "LL6", "processingTimeMs": 15, "limit": 5, "offset": 0, "estimatedTotalHits": 42, "facetDistribution": { "mass": { "110000": 1, "11500": 1, "1161": 1, "12000": 1, "1215.5": 1, "127000": 1, "15000": 1, "1676": 1, "1700": 1, "1710.5": 1, "18000": 1, "19000": 1, "220000": 1, "2220": 1, "22300": 1, "25000": 2, "265": 1, "271000": 1, "2840": 1, "3.3": 1, "3000": 1, "303": 1, "32000": 1, "34000": 1, "36.1": 1, "45000": 1, "460": 1, "478": 1, "483": 1, "5500": 2, "600": 1, "6000": 1, "67.8": 1, "678": 1, "680.5": 1, "6930": 1, "8": 1, "8300": 1, "840": 1, "8400": 1 }, "recclass": { "L/LL6": 3, "LL6": 39 } }, "facetStats": { "mass": { "min": 3.3, "max": 271000.0 } } } ``` </details> ## PR checklist Please check if your PR fulfills the following requirements: - [ ] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [ ] Have you read the contributing guidelines? - [ ] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-02-22 13:06:43 +00:00
curquiza	3eb9a08b5c	Update comments in version bump CI	2023-02-21 19:14:59 +01:00
ManyTheFish	900bae3d9d	keep phrases that has at least one word	2023-02-21 18:16:51 +01:00
ManyTheFish	28b7d73d4a	Remove an unefficient part of a test on milli	2023-02-21 18:16:51 +01:00
ManyTheFish	6841f167b4	Add test	2023-02-21 18:02:52 +01:00
bors[bot]	c88b6f331f	Merge #3482 3482: Optimize meilisearch uffizzi build r=curquiza a=waveywaves # Pull Request ## Related issue Fixes https://github.com/meilisearch/meilisearch/issues/3476 ## What does this PR do? even though docker cache was being used earlier for uffizzi builds, seems like the cache layers weren't persisting. This commit adds changes to move meilisearch building outside the dockerfile so that we can use the rust cache action. We are also building to the musl target so that the binary for meilisearch which is created can be used for the uffizzi ttyd image which uses alpine. Meilisearch build time brought to 5 mins example https://github.com/waveywaves/meilisearch/actions/runs/4142776058 we also update the version of uffizzi action used here which fixes another uffizzi bug where the environments are not deployed. https://app.uffizzi.com/github.com/waveywaves/meilisearch/pull/2 was built as a part of a test for this PR and we can be sure that the deployment works well now. ## PR checklist Please check if your PR fulfills the following requirements: - [ ] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [ ] Have you read the contributing guidelines? - [ ] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Vibhav Bobade <vibhav.bobde@gmail.com>	2023-02-21 16:06:53 +00:00
Vibhav Bobade	09a94e0db3	optimize meilisearch uffizzi build even though docker cache was being used earlier for uffizzi builds, seems like the cache layers weren't persisting. This commit adds changes to move meilisearch building outside the dockerfile so that we can use the rust cache action. We are also building to the musl target so that the binary for meilisearch which is created can be used for the uffizzi ttyd image which uses alpine.	2023-02-21 17:25:28 +05:30
bors[bot]	39407885c2	Merge #3347 3347: Enhance language detection r=irevoire a=ManyTheFish ## Summary Some completely unrelated Languages can share the same characters, in Meilisearch we detect the Languages using `whatlang`, which works well on large texts but fails on small search queries leading to a bad segmentation and normalization of the query. This PR now stores the Languages detected during the indexing in order to reduce the Languages list that can be detected during the search. ## Detail - Create a 19th database mapping the scripts and the Languages detected with the documents where the Language is detected - Fill the newly created database during indexing - Create an allow-list with this database and pass it to Charabia - Add a test ensuring that a Japanese request containing kanjis only is detected as Japanese and not Chinese ## Related issues Fixes #2403 Fixes #3513 Co-authored-by: f3r10 <frledesma@outlook.com> Co-authored-by: ManyTheFish <many@meilisearch.com> Co-authored-by: Many the fish <many@meilisearch.com>	2023-02-21 10:52:13 +00:00
bors[bot]	a3e41ba33e	Merge #3496 3496: Fix metrics feature r=irevoire a=james-2001 # Pull Request ## Related issue Resolves: #3469 See also: #2763 ## What does this PR do? As reported the metrics feature was broken by still using and old reference to `meilisearch_auth::actions`. This commit switches to the new location, `meilisearch_types::keys::actions`. The original issue was not that clear as to exactly what was broken, and the build logs have disappeared, but it seemed to just be this one line fix. If this is not the case and I've missed the mark let me know, and i'll head back to the drawing board. ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Co-authored-by: James <james.a.may.2001@gmail.com>	2023-02-21 10:13:11 +00:00
James	ce807d760b	Fix formatting issue on Opt struct tab in enable_metrics_route to fix cargo fmt issues Resolves: #3469 See also: #2763	2023-02-21 09:45:18 +00:00
ManyTheFish	bbecab8948	fix clippy	2023-02-21 10:18:44 +01:00
James	5cff435bf6	Add feature flags to Opt structure Resolves: #3469 See also: #2763	2023-02-21 07:41:41 +00:00
ManyTheFish	8aa808d51b	Merge branch 'main' into enhance-language-detection	2023-02-20 18:14:34 +01:00
bors[bot]	1e9ac00800	Merge #3505 3505: Csv delimiter r=irevoire a=irevoire Fixes https://github.com/meilisearch/meilisearch/issues/3442 Closes https://github.com/meilisearch/meilisearch/pull/2803 Specified in https://github.com/meilisearch/specifications/pull/221 This PR is a reimplementation of https://github.com/meilisearch/meilisearch/pull/2803, on the new engine. Thanks for your idea and initial PR `@MixusMinimax;` sorry I couldn’t update/merge your PR. Way too many changes happened on the engine in the meantime. Attention to reviewer; I had to update deserr to implement the support of deserializing `char`s ------- It introduces four new error messages; - Invalid value in parameter csvDelimiter: expected a string of one character, but found an empty string - Invalid value in parameter csvDelimiter: expected a string of one character, but found the following string of 5 characters: doggo - csv delimiter must be an ascii character. Found: 🍰 - The Content-Type application/json does not support the use of a csv delimiter. The csv delimiter can only be used with the Content-Type text/csv. And one error code; - `invalid_index_csv_delimiter` The `invalid_content_type` error code is now also used when we encounter the `csvDelimiter` query parameter with a non-csv content type. Co-authored-by: Tamo <tamo@meilisearch.com>	2023-02-20 17:01:36 +00:00
bors[bot]	b08a49a16e	Merge #3319 #3470 3319: Transparently resize indexes on MaxDatabaseSizeReached errors r=Kerollmops a=dureuill # Pull Request ## Related issue Related to https://github.com/meilisearch/meilisearch/discussions/3280, depends on https://github.com/meilisearch/milli/pull/760 ## What does this PR do? ### User standpoint - Meilisearch no longer fails tasks that encounter the `milli::UserError(MaxDatabaseSizeReached)` error. - Instead, these tasks are retried after increasing the maximum size allocated to the index where the failure occurred. ### Implementation standpoint - Add `Batch::index_uid` to get the `index_uid` of a batch of task if there is one - `IndexMapper::create_or_open_index` now takes an additional `size` argument that allows to (re)open indexes with a size different from the base `IndexScheduler::index_size` field - `IndexScheduler::tick` now returns a `Result<TickOutcome>` instead of a `Result<usize>`. This offers more explicit control over what the behavior should be wrt the next tick. - Add `IndexStatus::BeingResized` that contains a handle that a thread can use to await for the resize operation to complete and the index to be available again. - Add `IndexMapper::resize_index` to increase the size of an index. - In `IndexScheduler::tick`, intercept task batches that failed due to `MaxDatabaseSizeReached` and resize the index that caused the error, then request a new tick that will eventually handle the still enqueued task. ## Testing the PR The following diff can be applied to this branch to make testing the PR easier: <details> ```diff diff --git a/index-scheduler/src/index_mapper.rs b/index-scheduler/src/index_mapper.rs index 553ab45a..022b2f00 100644 --- a/index-scheduler/src/index_mapper.rs +++ b/index-scheduler/src/index_mapper.rs `@@` -228,13 +228,15 `@@` impl IndexMapper { drop(lock); + std:🧵:sleep_ms(2000); + let current_size = index.map_size()?; let closing_event = index.prepare_for_closing(); - log::info!("Resizing index {} from {} to {} bytes", name, current_size, current_size * 2); + log::error!("Resizing index {} from {} to {} bytes", name, current_size, current_size * 2); closing_event.wait(); - log::info!("Resized index {} from {} to {} bytes", name, current_size, current_size * 2); + log::error!("Resized index {} from {} to {} bytes", name, current_size, current_size * 2); let index_path = self.base_path.join(uuid.to_string()); let index = self.create_or_open_index(&index_path, None, 2 * current_size)?; `@@` -268,8 +270,10 `@@` impl IndexMapper { match index { Some(Available(index)) => break index, Some(BeingResized(ref resize_operation)) => { + log::error!("waiting for resize end"); // Deadlock: no lock taken while doing this operation. resize_operation.wait(); + log::error!("trying our luck again!"); continue; } Some(BeingDeleted) => return Err(Error::IndexNotFound(name.to_string())), diff --git a/index-scheduler/src/lib.rs b/index-scheduler/src/lib.rs index 11b17d05..242dc095 100644 --- a/index-scheduler/src/lib.rs +++ b/index-scheduler/src/lib.rs `@@` -908,6 +908,7 `@@` impl IndexScheduler { /// /// Returns the number of processed tasks. fn tick(&self) -> Result<TickOutcome> { + log::error!("ticking!"); #[cfg(test)] { *self.run_loop_iteration.write().unwrap() += 1; diff --git a/meilisearch/src/main.rs b/meilisearch/src/main.rs index 050c825a..63f312f6 100644 --- a/meilisearch/src/main.rs +++ b/meilisearch/src/main.rs `@@` -25,7 +25,7 `@@` fn setup(opt: &Opt) -> anyhow::Result<()> { #[actix_web::main] async fn main() -> anyhow::Result<()> { - let (opt, config_read_from) = Opt::try_build()?; + let (mut opt, config_read_from) = Opt::try_build()?; setup(&opt)?; `@@` -56,6 +56,8 `@@` We generated a secure master key for you (you can safely copy this token): _ => (), } + opt.max_index_size = byte_unit::Byte::from_str("1MB").unwrap(); + let (index_scheduler, auth_controller) = setup_meilisearch(&opt)?; #[cfg(all(not(debug_assertions), feature = "analytics"))] ``` </details> Mainly, these debug changes do the following: - Set the default index size to 1MiB so that index resizes are initially frequent - Turn some logs from info to error so that they can be displayed with `--log-level ERROR` (hiding the other infos) - Add a long sleep between the beginning and the end of the resize so that we can observe the `BeingResized` index status (otherwise it would never come up in my tests) ## Open questions - Is the growth factor of x2 the correct solution? For a `Vec` in memory it makes sense, but here we're manipulating quantities that are potentially in the order of 500GiBs. For bigger indexes it may make more sense to add at most e.g. 100GiB on each resize operation, avoiding big steps like 500GiB -> 1TiB. ## PR checklist Please check if your PR fulfills the following requirements: - [ ] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [ ] Have you read the contributing guidelines? - [ ] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! 3470: Autobatch addition and deletion r=irevoire a=irevoire This PR adds the capability to meilisearch to batch document addition and deletion together. Fix https://github.com/meilisearch/meilisearch/issues/3440 -------------- Things to check before merging; - [x] What happens if we delete multiple time the same documents -> add a test - [x] If a documentDeletion gets batched with a documentAddition but the index doesn't exist yet? It should not work Co-authored-by: Louis Dureuil <louis@meilisearch.com> Co-authored-by: Tamo <tamo@meilisearch.com>	2023-02-20 15:00:19 +00:00
ManyTheFish	23f4e82b53	Add test ensuring that Meilisearch works on kanji only requests	2023-02-20 15:43:29 +01:00
Many the fish	119e6d8811	Update milli/src/search/mod.rs Co-authored-by: Tamo <tamo@meilisearch.com>	2023-02-20 15:33:10 +01:00
bors[bot]	a8f6f108e0	Merge #3515 3515: Consider null as a valid geo field r=irevoire a=irevoire Fix #3497 Associated spec; https://github.com/meilisearch/specifications/pull/222 Co-authored-by: Tamo <tamo@meilisearch.com>	2023-02-20 14:12:55 +00:00
Tamo	1479050f7a	apply review suggestions	2023-02-20 14:53:37 +01:00
bors[bot]	97b8c32e22	Merge #3514 3514: Bump version of mini-dashboard to v0.2.6 r=irevoire a=bidoubiwa Update the version of the mini-dashboard to v0.2.6. See [release notes](https://github.com/meilisearch/mini-dashboard/releases/tag/v0.2.6). Co-authored-by: Charlotte Vermandel <charlottevermandel@gmail.com>	2023-02-20 13:21:00 +00:00
ManyTheFish	cb8d5f2d4b	Update Charabia to 0.7.1	2023-02-20 14:00:31 +01:00
Louis Dureuil	35f6c624bc	Make sure we don't leave the in memory hashmap in an inconsistent state	2023-02-20 13:55:32 +01:00
Louis Dureuil	1116788475	Resize indexes when they're full	2023-02-20 13:55:32 +01:00
Louis Dureuil	951a5b5832	Add IndexMapper::resize_index fn	2023-02-20 13:55:32 +01:00
Louis Dureuil	1c670d7fa0	Add IndexStatus::BeingResized	2023-02-20 13:55:32 +01:00
Louis Dureuil	6cc3797aa1	IndexScheduler::tick returns a TickOutcome	2023-02-20 13:55:31 +01:00
Louis Dureuil	faf1e17a27	`create_or_open_index` takes a `map_size` argument	2023-02-20 13:55:31 +01:00
Louis Dureuil	4c519c2ab3	Add Batch::index_uid	2023-02-20 13:55:31 +01:00
Louis Dureuil	eb28d4c525	add facet test	2023-02-20 13:52:28 +01:00
Louis Dureuil	9ac981d025	Remove some clippy type complexity warns by deboxing iters	2023-02-20 13:52:27 +01:00
Louis Dureuil	74859ecd61	Add min and max facet stats	2023-02-20 13:52:27 +01:00
Louis Dureuil	8ae441a4db	Update usage of iterators	2023-02-20 13:52:27 +01:00
Louis Dureuil	042d86cbb3	facet sort ascending/descending now also return the values	2023-02-20 13:52:27 +01:00
Charlotte Vermandel	dd120e0e16	Bump version of mini-dashboard to v0.2.6	2023-02-20 13:45:57 +01:00
Tamo	18796d6e6a	Consider null as a valid geo object	2023-02-20 13:45:51 +01:00
bors[bot]	c91bfeaf15	Merge #3467 3467: Identify builds git tagged with `prototype-...` in CLI and analytics r=curquiza a=dureuill # Pull Request ## What does this PR do? - Parses the last git tag to extract a prototype name if: - Current build uses the prototype tag (not after the tag) precisely - The prototype tag name respects the following conditions: 1. starts with `prototype-` 2. ends with a number 3. the hyphen-separated segment right before the number is not a number (required to reject commits after the tag). - Display the prototype name in the launch summary in the CLI - Send the prototype name to analytics if any - Update prototypes instructions in CONTRIBUTING.md \|`VERGEN_GIT_SEMVER_LIGHTWEIGHT` value \| Prototype \| \|---\|---\| \| `Some("prototype-geo-bounding-box-0-139-gcde89018")` \| `None` (does not end with a number) \| \| `Some("prototype-geo-bounding-box-0-139-89018")` \| `None` (before the last segment is a number) \| \| `Some("prototype-geo-bounding-box-0")` \| `Some("prototype-geo-bounding-box-0")` \| \| `Some("prototype-geo-bounding-box")` \| `None` (does not end with a number") \| \| `Some("geo-bounding-box-0")` \| `None` (does not start with "prototype") \| \| `None` \| `None` \| Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-02-20 09:27:51 +00:00
James	91048d209d	Fix metrics feature Metrics feature was relying on old references. Refactored with inspiration from the `get_stats` method in `meilisearch/src/routes/lib.rs`. `enable_metrics_routes` added to options in `segment_analytics`. Resolves: #3469 See also: #2763	2023-02-17 20:11:57 +00:00
bors[bot]	28961b2ad1	Merge #3499 3499: Use the workspace inheritance r=Kerollmops a=irevoire Use the workspace inheritance [introduced in rust 1.64](https://blog.rust-lang.org/2022/09/22/Rust-1.64.0.html#cargo-improvements-workspace-inheritance-and-multi-target-builds). It allows us to define the version of meilisearch once in the main `Cargo.toml` and let all the other `Cargo.toml` uses this version. `@curquiza` I added you as a reviewer because I had to patch some CI scripts And `@Kerollmops,` I had to bump the `cargo_toml` crates because our version was getting old and didn't support the feature yet. Also, in another PR, I would like to unify some of our dependencies to ensure we always stay in sync between all our crates. Co-authored-by: Tamo <tamo@meilisearch.com>	2023-02-17 09:52:29 +00:00
Tamo	895ab2906c	apply review suggestions	2023-02-16 18:42:47 +01:00
Tamo	f11c7d4b62	cargo run execute meilisearch by default	2023-02-16 18:03:45 +01:00
Tamo	e79f6f87f6	make cargo fmt&clippy happy	2023-02-16 18:00:40 +01:00
Tamo	5367d8f05a	add two tests on the indexing of csvs	2023-02-16 17:37:11 +01:00
Tamo	52686da028	test various error on the document ressource	2023-02-16 17:37:10 +01:00
Tamo	8c074f5028	implements the csv delimiter without tests Co-authored-by: Maxi Barmetler <maxi.barmetler@gmail.com>	2023-02-16 17:35:36 +01:00
Louis Dureuil	49e18da23e	Do not escape tag name $() syntax is not interpreted by the Dockerfile	2023-02-16 10:53:14 +01:00
Louis Dureuil	54240db495	Add note in code so one does not forget next time	2023-02-16 10:53:14 +01:00
Louis Dureuil	e1ed4bc750	Change Dockerfile to also pass the VERGEN_GIT_SEMVER_LIGHTWEIGHT when building	2023-02-16 10:53:14 +01:00
Louis Dureuil	9bd1cfb3a3	Ignore -dirty flag	2023-02-16 10:53:14 +01:00
Louis Dureuil	a341c94871	Update contributing.md	2023-02-16 10:53:14 +01:00
Louis Dureuil	f46cf46b8c	Add prototype to analytics if any	2023-02-16 10:53:14 +01:00
Louis Dureuil	c3a30a5a91	If using a prototype, display its name at Meilisearch startup	2023-02-16 10:53:14 +01:00
bors[bot]	143e3cf948	Merge #3490 3490: Fix attributes set candidates r=curquiza a=ManyTheFish # Pull Request Fix attributes set candidates for v1.1.0 ## details The attribute criterion was not returning the remaining candidates when its internal algorithm was been exhausted. We had a loss of candidates by the attribute criterion leading to the bug reported in the issue linked below. After some investigation, it seems that it was the only criterion that had this behavior. We are now returning the remaining candidates instead of an empty bitmap. ## Related issue Fixes #3483 PR on milli for v1.0.1: https://github.com/meilisearch/milli/pull/777 Co-authored-by: ManyTheFish <many@meilisearch.com>	2023-02-15 17:38:07 +00:00
Tamo	ab2adba183	update our CI scripts accordingly	2023-02-15 13:56:24 +01:00
Tamo	74d1a67a99	Use the workspace inheritance feature of rust 1.64	2023-02-15 13:51:07 +01:00
bors[bot]	91ce8a5e67	Merge #3492 3492: Bump deserr r=Kerollmops a=irevoire Bump deserr to the latest version; - We now use the default actix-web extractors that deserr provides (which were copy/pasted from meilisearch) - We also use the default `JsonError` message provided by deserr instead of defining our own in meilisearch - Finally, we get the new `did you mean?` error message. Fix #3493 Co-authored-by: Tamo <tamo@meilisearch.com>	2023-02-15 10:05:05 +00:00
bors[bot]	fd7ae1883b	Merge #3495 3495: Add tests with rust nightly in CI r=curquiza a=ztkmkoo # Pull Request ## Related issue Fixes #3402 ## What does this PR do? - add ci test with rust nightly - make test with rust stable not run on schedule event ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Kebron <ztkmkoo@gmail.com>	2023-02-15 07:53:17 +00:00
Tamo	42a3cdca66	get rids of the unwrap_any function in favor of take_cf_content	2023-02-14 20:06:31 +01:00
Tamo	a43765d454	use the pre-defined deserr extractors	2023-02-14 20:05:30 +01:00
Tamo	769576fd94	get rids of the whole error_message module since it has been integrated into the last version of deserr	2023-02-14 20:05:27 +01:00
Tamo	8fb7b1d10f	bump deserr	2023-02-14 20:04:30 +01:00
bors[bot]	d494c29768	Merge #3479 3479: Unify "Bad latitude" & "Bad longitude" errors r=irevoire a=cymruu # Pull Request ## Related issue Fix part of #3006 ## What does this PR do? - Moved out `BadGeoLat`, `BadGeoLng`, `BadGeoBoundingBoxTopIsBelowBottom` from `FilterError` into newly introduced error type `ParseGeoError`. - Renamed `BadGeo` error to `ReservedGeo` - Used new `ParseGeoError` type in `FilterError` and `AscDescError` Screenshot: ![image](https://user-images.githubusercontent.com/2981598/217927231-fe23b6a3-2ea8-4145-98af-38eb61c4ff16.png) I ran `cargo test --package milli -- --test-threads 1` and tests passed. `--test-threads` was set to 1 because my OS complained about too many opened files. ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [ ] Have you made sure that the title is accurate and descriptive of the changes? Co-authored-by: Filip Bachul <filipbachul@gmail.com> Co-authored-by: filip <filipbachul@gmail.com>	2023-02-14 18:35:51 +00:00
Tamo	74dcfe9676	Fix a bug when you update a document that was already present in the db, deleted and then inserted again in the same transform	2023-02-14 19:09:40 +01:00
Tamo	1b1703a609	make a small optimization to merge obkvs a little bit faster	2023-02-14 18:32:41 +01:00
James	62358bd31c	Fix metrics feature As reported the metrics feature was broken by still using and old reference to `meilisearch_auth::actions`. This commit switches to the new location, `meilisearch_types::keys::actions`. Resolves: #3469 See also: #2763	2023-02-14 17:29:38 +00:00
Tamo	fb5e4957a6	fix and test the early exit in case a grenad ends with a deletion	2023-02-14 18:23:57 +01:00
Tamo	8de3c9f737	Update milli/src/update/index_documents/transform.rs Co-authored-by: Clément Renault <clement@meilisearch.com>	2023-02-14 17:57:14 +01:00
Tamo	43a19d0709	document the operation enum + the grenads	2023-02-14 17:55:26 +01:00
Tamo	29d14bed90	get rids of the let/else syntax	2023-02-14 17:45:46 +01:00
bors[bot]	f3b54337f9	Merge #3174 3174: Allow wildcards at the end of index names for API Keys and Tenant tokens r=irevoire a=Kerollmops This PR introduces the wildcards at the end of the index names when identifying indexes in the API Keys and tenant tokens. It fixes #2788 and fixes #2908. This PR is based on `@akhildevelops'` work. Note that when a tenant token filter is chosen to restrict a search, it is always the most restrictive pattern that is chosen. If we have an index pattern _prod_ that defines _filter1_ and _p_ that defines _filter2_, the engine will choose _filter1_ over _filter2_ as it is defined for a most restrictive pattern, _prod_. This restrictiveness is defined by 1. is it exact, without __ 2. the length of the pattern. It is a continuation of work that has already started and should close #2869. Co-authored-by: Clément Renault <clement@meilisearch.com> Co-authored-by: Kerollmops <clement@meilisearch.com>	2023-02-14 16:12:01 +00:00
Clément Renault	7f3ae40204	Remove a useless comment regarding the index pattern error code	2023-02-14 17:09:20 +01:00
Filip Bachul	a53536836b	fmt	2023-02-14 17:04:22 +01:00
Kebron	b095325bf8	Add tests with rust nightly in CI	2023-02-14 15:33:12 +00:00
Filip Bachul	d7ad39ad77	fix: clippy error	2023-02-14 00:15:35 +01:00
Filip Bachul	849de089d2	add thiserror for AscDescError	2023-02-14 00:15:35 +01:00
filip	7f25007d31	Update milli/src/asc_desc.rs Co-authored-by: Tamo <irevoire@protonmail.ch>	2023-02-14 00:15:35 +01:00
Filip Bachul	c810af3ebf	implement From<ParseGeoError> for AscDescError	2023-02-14 00:15:35 +01:00
Filip Bachul	c0b77773ba	fmt asc_desc	2023-02-14 00:15:35 +01:00
Filip Bachul	7481559e8b	move BadGeo to FilterError	2023-02-14 00:15:35 +01:00
Filip Bachul	83c765ce6c	implement From<ParseGeoError> for FilterError	2023-02-14 00:15:35 +01:00
Filip Bachul	4c91037602	use ParseGeoError in sort parser	2023-02-14 00:15:35 +01:00
Filip Bachul	825923f6fc	export ParseGeoError	2023-02-14 00:15:35 +01:00
Filip Bachul	e405702733	chore: introduce new error ParseGeoError type	2023-02-14 00:15:35 +01:00
ManyTheFish	6fa877efb0	Fix attributes set candidates	2023-02-13 17:49:52 +01:00
Kerollmops	4b1cd10653	Return an internal error when index pattern should be valid	2023-02-13 17:49:42 +01:00
Clément Renault	47748395dc	Update an authentication comment Co-authored-by: Many the fish <many@meilisearch.com>	2023-02-13 17:20:08 +01:00
bors[bot]	ff595156d7	Merge #3480 3480: Gitignore vscode & jetbrains IDE folders r=curquiza a=AymanHamdoun # Pull Request ## Related issue There is no issue for it, and i couldn't find an appropriate category to make an issue for it. ## What does this PR do? - Its just a gitignore edit so people who use vscode and jetbrains IDEs (like IntelliJ) dont have to deal with committing the folder the IDE generates to store local project configs by mistake. (I honestly wanted to fork the repo to add something else but this bothered me enough to make a PR for it first) ## PR checklist Please check if your PR fulfills the following requirements: - [ ✔️ ] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [✔️ ] Have you read the contributing guidelines? - [✔️ ] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Ayman <ayman.s.hamdoun@gmail.com>	2023-02-13 10:47:25 +00:00
Ayman	8770088df3	remove idea folder	2023-02-10 11:45:02 +04:00
Ayman	827c1c8447	edit gitignore to ignore .idea and .vscode folders	2023-02-10 11:42:19 +04:00
Clément Renault	764df24b7d	Make clippy happy (again)	2023-02-09 13:21:20 +01:00
Clément Renault	4570d5bf3a	Merge remote-tracking branch 'origin/main' into temp-wildcard	2023-02-09 13:14:05 +01:00
Tamo	746b31c1ce	makes clippy happy	2023-02-09 12:23:01 +01:00
Tamo	eaad84bd1d	fix the test to handle the document deletion correctly	2023-02-09 11:29:13 +01:00
Kerollmops	c690c4fec4	Added and modified the current API Key and Tenant Token tests	2023-02-09 11:17:30 +01:00
Tamo	ea9ac46f28	stop autobatching the deletion without the index creation right with the addition	2023-02-08 21:24:27 +01:00
Tamo	93db755d57	add a test to ensure we handle correctly a deletion of multiple time the same document	2023-02-08 21:03:34 +01:00
Tamo	93f130a400	fix all warnings	2023-02-08 20:57:35 +01:00
Tamo	860c993ef7	Handle the autobatching of deletion and addition in the scheduler	2023-02-08 20:53:19 +01:00
Tamo	67dda0678f	cleanup the autobatcher a little bit	2023-02-08 18:10:59 +01:00
Tamo	2db6347686	update the autobatcher to batch the addition and deletion together	2023-02-08 18:07:59 +01:00
Tamo	421a9cf05e	provide a new method on the transform to remove documents	2023-02-08 16:06:09 +01:00
Kerollmops	7b4b57ecc8	Fix the current tests	2023-02-08 14:54:05 +01:00
Tamo	8f64fba1ce	rewrite the current transform to handle a new byte specifying the kind of operation it's merging	2023-02-08 12:53:38 +01:00
bors[bot]	9882029fa4	Merge #3456 3456: Bump tokio from 1.24.1 to 1.24.2 r=curquiza a=dependabot[bot] Bumps [tokio](https://github.com/tokio-rs/tokio) from 1.24.1 to 1.24.2. <details> <summary>Commits</summary> <ul> <li>See full diff in <a href="https://github.com/tokio-rs/tokio/commits">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=tokio&package-manager=cargo&previous-version=1.24.1&new-version=1.24.2)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) Dependabot will resolve any conflicts with this PR as long as you don't alter it yourself. You can also trigger a rebase manually by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) You can disable automated security fix PRs for this repo from the [Security Alerts page](https://github.com/meilisearch/meilisearch/network/alerts). </details> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2023-02-07 13:28:42 +00:00
dependabot[bot]	5f56e6dd58	Bump tokio from 1.24.1 to 1.24.2 Bumps [tokio](https://github.com/tokio-rs/tokio) from 1.24.1 to 1.24.2. - [Release notes](https://github.com/tokio-rs/tokio/releases) - [Commits](https://github.com/tokio-rs/tokio/commits) --- updated-dependencies: - dependency-name: tokio dependency-type: direct:production ... Signed-off-by: dependabot[bot] <support@github.com>	2023-02-07 12:14:05 +00:00
bors[bot]	c88c3637b4	Merge #3461 3461: Bring v1 changes into main r=curquiza a=Kerollmops Also bring back changes in milli (the remote repository) into main done during the pre-release Co-authored-by: Loïc Lecrenier <loic.lecrenier@me.com> Co-authored-by: bors[bot] <26634292+bors[bot]@users.noreply.github.com> Co-authored-by: curquiza <curquiza@users.noreply.github.com> Co-authored-by: Tamo <tamo@meilisearch.com> Co-authored-by: Philipp Ahlner <philipp@ahlner.com> Co-authored-by: Kerollmops <clement@meilisearch.com>	2023-02-07 11:27:27 +00:00
bors[bot]	97fd9ac493	Merge #3405 3405: Implement geo bounding box r=irevoire a=curquiza Following https://github.com/meilisearch/milli/pull/672 (work from `@gmourier)` Fixes #2761 Co-authored-by: Guillaume Mourier <guillaume@meilisearch.com> Co-authored-by: Louis Dureuil <louis@meilisearch.com> Co-authored-by: Tamo <tamo@meilisearch.com>	2023-02-07 09:55:20 +00:00
bors[bot]	821d92b5d0	Merge #3407 3407: Add Cargo feature for LMDB's POSIX semaphores r=dureuill a=GregoryConrad See https://github.com/meilisearch/milli/pull/757 Co-authored-by: Gregory Conrad <gregorysconrad@gmail.com>	2023-02-07 08:25:20 +00:00
bors[bot]	0b60928cbc	Merge #3199 3199: Fixup dumps-destination -> dump-directory section header in help link r=curquiza a=dureuill # Pull Request ## Related issue See https://github.com/meilisearch/product/discussions/560#discussioncomment-4323938 ## What does this PR do? - change link in help message to the future new section header #dump-directory Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-02-06 17:49:32 +00:00
Tamo	42114325cd	Apply suggestions from code review Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-02-06 18:07:00 +01:00
Tamo	7a38fe624f	throw an error if the top left corner is found below the bottom right corner	2023-02-06 17:50:47 +01:00
Tamo	1b005f697d	update the syntax of the geoboundingbox filter to uses brackets instead of parens around lat and lng	2023-02-06 16:50:27 +01:00
Kerollmops	fbec48f56e	Merge remote-tracking branch 'milli/main' into bring-v1-changes	2023-02-06 16:48:10 +01:00
Kerollmops	a377a49218	Make meiliserach depend on the local milli	2023-02-06 16:44:43 +01:00
Kerollmops	41cbaad1cb	Revert "Add git config about ownershio in Docker CI" This reverts commit `e269027cdd`.	2023-02-06 16:42:16 +01:00
Kerollmops	a015e232ab	Merge remote-tracking branch 'origin/release-v1.0.0' into bring-v1-changes	2023-02-06 16:41:10 +01:00
Tamo	3ebc99473f	Apply suggestions from code review Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-02-06 13:29:37 +01:00
bors[bot]	fadea504ed	Merge #3451 3451: Pin Rust version in Clippy job r=dureuill a=curquiza Avoid "surprising" CI failure because of clippy when rust is releasing a new version Co-authored-by: curquiza <clementine@meilisearch.com> Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com>	2023-02-06 12:19:35 +00:00
Tamo	d27007005e	comments the geoboundingbox + forbid the usage of the lexeme method which could introduce bugs	2023-02-06 11:36:49 +01:00
Tamo	fcb09ccc3d	add tests on the geoBoundingBox	2023-02-02 18:19:56 +01:00
bors[bot]	734a9ecea8	Merge #3040 3040: feat: create a preview environment for every PR using Uffizzi r=curquiza a=waveywaves # Pull Request ## Related discussion (was created as an issue initially) https://github.com/meilisearch/meilisearch/discussions/2883 ## What does this PR do? This PR adds gha workflows to create preview environments on every PR. This workflow also posts the preview url as a comment on the PR. [This PR created against my fork of meilisearch](https://github.com/waveywaves/meilisearch/pull/2) demonstrates how this change behaves. In [the demo preview](https://pr-2-deployment-7396-meilisearch.app.uffizzi.com/) you can run the `meilisearch` binary built from the PR and can access meilisearch running from the PR by adding `/meilisearch` to the url of the PR. eg: I go to the demo preview at the URL https://app.uffizzi.com/github.com/waveywaves/meilisearch/pull/2, run `meilisearch` in the terminal. I can access this running instance of `meilisearch` in the preview env fromhttps://pr-2-deployment-7396-meilisearch.app.uffizzi.com/meilisearch ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [ ] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Vibhav Bobade <vibhav.bobde@gmail.com>	2023-02-02 16:06:38 +00:00
curquiza	69fcd3d05e	Add comment information about the cron job	2023-02-02 15:58:03 +01:00
Clémentine Urquizar - curqui	1ca7778e6a	Update .github/workflows/create-issue-dependencies.yml Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-02-02 15:54:33 +01:00
curquiza	a11d992923	Update issue description for the dependency updates	2023-02-02 15:33:38 +01:00
curquiza	781691191a	Pin Rust version in Clippy job	2023-02-02 15:22:58 +01:00
Louis Dureuil	ae8660e585	Add Token::original_span rather than making Token::span pub	2023-02-02 15:03:34 +01:00
Guillaume Mourier	d80ce00623	Update insta test	2023-02-02 12:34:51 +01:00
Guillaume Mourier	2d66fdc8e9	Apply review comments	2023-02-02 12:34:51 +01:00
Guillaume Mourier	b297b5deb0	cargo fmt	2023-02-02 12:34:49 +01:00
Guillaume Mourier	0d71c80ba6	add tests	2023-02-02 12:31:27 +01:00
Guillaume Mourier	b2054d3f6c	Add insta test on geo filters whitespacing	2023-02-02 12:27:58 +01:00
Guillaume Mourier	65a3086cf1	fix test	2023-02-02 12:27:58 +01:00
Guillaume Mourier	426d63b01b	Update insta test suite	2023-02-02 12:27:56 +01:00
Guillaume Mourier	b078477d80	Add error handling and earth lap collision with bounding box	2023-02-02 12:17:38 +01:00
Guillaume Mourier	5c525168a0	Add _geoBoundingBox parser	2023-02-02 11:57:21 +01:00
bors[bot]	39b62b7158	Merge #3436 3436: Add more detailed contribution instructions for tests r=irevoire a=dureuill Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-02-02 10:19:41 +00:00
bors[bot]	3f97f630ed	Merge #3448 3448: Bump docker/build-push-action from 3 to 4 r=curquiza a=dependabot[bot] Bumps [docker/build-push-action](https://github.com/docker/build-push-action) from 3 to 4. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/docker/build-push-action/releases">docker/build-push-action's releases</a>.</em></p> <blockquote> <h2>v4.0.0</h2> <blockquote> <p><strong>Note</strong></p> <p>Buildx v0.10 enables support for a minimal <a href="https://slsa.dev/provenance/">SLSA Provenance</a> attestation, which requires support for <a href="https://github.com/opencontainers/image-spec">OCI-compliant</a> multi-platform images. This may introduce issues with registry and runtime support (e.g. <a href="https://github-redirect.dependabot.com/docker/buildx/issues/1533">Google Cloud Run and AWS Lambda</a>). You can optionally disable the default provenance attestation functionality using <code>provenance: false</code>.</p> </blockquote> <ul> <li>Revert disable provenance by default if not set by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` in <a href="https://github-redirect.dependabot.com/docker/build-push-action/pull/784">docker/build-push-action#784</a></li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/build-push-action/compare/v3.3.1...v4.0.0">https://github.com/docker/build-push-action/compare/v3.3.1...v4.0.0</a></p> <h2>v3.3.1</h2> <ul> <li>Disable provenance by default if not set by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/781">#781</a>)</li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/build-push-action/compare/v3.3.0...v3.3.1">https://github.com/docker/build-push-action/compare/v3.3.0...v3.3.1</a></p> <h2>v3.3.0</h2> <blockquote> <p><strong>Note</strong></p> <p>Buildx v0.10 enables support for a minimal <a href="https://slsa.dev/provenance/">SLSA Provenance</a> attestation, which requires support for <a href="https://github.com/opencontainers/image-spec">OCI-compliant</a> multi-platform images. This may introduce issues with registry and runtime support (e.g. <a href="https://github-redirect.dependabot.com/docker/buildx/issues/1533">Google Cloud Run and AWS Lambda</a>). You can optionally disable the default provenance attestation functionality using <code>provenance: false</code>.</p> </blockquote> <ul> <li>Add <code>attests</code>, <code>provenance</code> and <code>sbom</code> inputs by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/746">#746</a> <a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/759">#759</a>)</li> <li>Log GitHub Actions runtime token access controls by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/707">#707</a>)</li> <li>Examples moved to <a href="https://docs.docker.com/build/ci/github-actions/examples/">docs website</a> by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/718">#718</a>)</li> <li>Bump minimatch from 3.0.4 to 3.1.2 (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/732">#732</a>)</li> <li>Bump csv-parse from 5.3.0 to 5.3.3 (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/729">#729</a>)</li> <li>Bump json5 from 2.2.0 to 2.2.3 (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/749">#749</a>)</li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/build-push-action/compare/v3.2.0...v3.3.0">https://github.com/docker/build-push-action/compare/v3.2.0...v3.3.0</a></p> <h2>v3.2.0</h2> <ul> <li>Remove workaround for <code>setOutput</code> by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/704">#704</a>)</li> <li>Docs: fix Git context link and add more details about subdir support by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/685">#685</a>)</li> <li>Docs: named context by <a href="https://github.com/baibaratsky"><code>`@baibaratsky</code></a>` and <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/665">#665</a>)</li> <li>Bump <code>`@actions/core</code>` from 1.9.0 to 1.10.0 (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/667">#667</a> <a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/695">#695</a>)</li> <li>Bump <code>`@actions/github</code>` from 5.0.3 to 5.1.1 (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/696">#696</a>)</li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/build-push-action/compare/v3.1.1...v3.2.0">https://github.com/docker/build-push-action/compare/v3.1.1...v3.2.0</a></p> <h2>v3.1.1</h2> <ul> <li>Fix GitHub token not passed with Git context if subdir defined by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/663">#663</a>)</li> <li>Replace deprecated <code>fs.rmdir</code> with <code>fs.rm</code> by <a href="https://github.com/bendrucker"><code>`@bendrucker</code></a>` (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/657">#657</a>)</li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/build-push-action/compare/v3.1.0...v3.1.1">https://github.com/docker/build-push-action/compare/v3.1.0...v3.1.1</a></p> <h2>v3.1.0</h2> <ul> <li><code>no-cache-filters</code> input by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/653">#653</a>)</li> <li>Bump <code>`@actions/github</code>` from 5.0.1 to 5.0.3 (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/619">#619</a>)</li> <li>Bump <code>`@actions/core</code>` from 1.6.0 to 1.9.0 (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/620">#620</a> <a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/637">#637</a>)</li> <li>Bump csv-parse from 5.0.4 to 5.3.0 (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/623">#623</a> <a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/650">#650</a>)</li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/build-push-action/compare/v3.0.0...v3.1.0">https://github.com/docker/build-push-action/compare/v3.0.0...v3.1.0</a></p> </blockquote> </details> <details> <summary>Commits</summary> <ul> <li><a href="`3b5e8027fc`"><code>3b5e802</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/784">#784</a> from crazy-max/enable-provenance</li> <li><a href="`02d3266a89`"><code>02d3266</code></a> update generated content</li> <li><a href="`f403dafe18`"><code>f403daf</code></a> revert disable provenance by default if not set</li> <li>See full diff in <a href="https://github.com/docker/build-push-action/compare/v3...v4">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=docker/build-push-action&package-manager=github_actions&previous-version=3&new-version=4)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) You can trigger a rebase of this PR by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) </details> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2023-02-01 18:13:15 +00:00
ManyTheFish	0bc1a18f52	Use Languages list detected during indexing at search time	2023-02-01 18:57:43 +01:00
ManyTheFish	643d99e0f9	Add expectancy test	2023-02-01 18:39:54 +01:00
Kerollmops	a36b1dbd70	Fix the tasks with the new patterns	2023-02-01 18:21:45 +01:00
dependabot[bot]	5672165e44	Bump docker/build-push-action from 3 to 4 Bumps [docker/build-push-action](https://github.com/docker/build-push-action) from 3 to 4. - [Release notes](https://github.com/docker/build-push-action/releases) - [Commits](https://github.com/docker/build-push-action/compare/v3...v4) --- updated-dependencies: - dependency-name: docker/build-push-action dependency-type: direct:production update-type: version-update:semver-major ... Signed-off-by: dependabot[bot] <support@github.com>	2023-02-01 17:02:17 +00:00
Kerollmops	d563ed8a39	Making it work with index uid patterns	2023-02-01 17:51:30 +01:00
bors[bot]	36cae3b480	Merge #3399 3399: Rework technical information in the README r=Kerollmops a=curquiza Following this https://github.com/meilisearch/meilisearch/pull/3346#discussion_r1073289399 Co-authored-by: curquiza <clementine@meilisearch.com> Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com>	2023-02-01 14:34:55 +00:00
ManyTheFish	064158e4e2	Update test	2023-02-01 15:34:01 +01:00
ManyTheFish	77d32d0ee8	Fix codec deserialization	2023-02-01 15:26:26 +01:00
ManyTheFish	f4569b04ad	Update Charabia version	2023-02-01 15:26:26 +01:00
bors[bot]	5e12af88e2	Merge #3445 3445: Bump milli to v0.41.1 r=curquiza a=dureuill # Pull Request ## Related issue Fixes #3438. ## What does this PR do? - Bump milli to [v0.41.1](https://github.com/meilisearch/milli/releases/tag/v0.41.1) that includes a bugfix for #3438 Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-02-01 11:07:46 +00:00
Louis Dureuil	231067a1c4	Bump milli to v0.41.1	2023-02-01 11:53:39 +01:00
Vibhav Bobade	2a1a7ef00a	Integrate Uffizzi	2023-02-01 13:06:27 +05:30
bors[bot]	758b4acea7	Merge #776 776: Reduce incremental indexing time of `words_prefix_position_docids` DB r=curquiza a=loiclec Fixes partially https://github.com/meilisearch/milli/issues/605 The `words_prefix_position_docids` can easily contain millions of entries. Thus, iterating over it can be very expensive. But we do so needlessly for every document addition tasks. It can sometimes cause indexing performance issues when : - a user sends many `documentAdditionOrUpdate` tasks that cannot be all batched together (for example if they are interspersed with `documentDeletion` tasks) - the documents contain long, diverse text fields, thus increasing the number of entries in `words_prefix_position_docids` - the index has accumulated many soft-deleted documents, further increasing the size of `words_prefix_position_docids` - the machine running Meilisearch does not have great IO performance (e.g. slow SSD, or quota-limited by the cloud provider) Note, before approving the PR: the only changed file should be `milli/src/update/words_prefix_position_docids.rs`. Co-authored-by: Loïc Lecrenier <loic.lecrenier@me.com>	2023-01-31 15:52:28 +00:00
bors[bot]	20f8184c06	Merge #3441 3441: Fix import of dump v2 r=dureuill a=irevoire # Pull Request This bug was introduced because of a mistake we did earlier: We said the last version to export dump v2 was the v0.21.0 while it was the v0.22.0. To fix the bug I updated our whole v2 reader to use the code from meilisearch v0.22.0. Also: - Import the bugged dump in the tests - Test the import of this dump in the v2 reader and current reader ## Related issue Fixes #3435 Co-authored-by: Tamo <tamo@meilisearch.com>	2023-01-31 13:23:57 +00:00
bors[bot]	2f8ebd0501	Merge #3439 3439: Add git config about ownership in Docker CI r=curquiza a=curquiza The docker CI si failing because of git usage: https://github.com/meilisearch/meilisearch/actions/runs/4053334082/jobs/6973827940 <img width="960" alt="Capture d’écran 2023-01-31 à 12 12 44" src="https://user-images.githubusercontent.com/20380692/215745119-b866bcf2-7077-48e4-b018-7a2085b23680.png"> > fatal: detected dubious ownership in repository at '/home/meili/actions-runner/_work/meilisearch/meilisearch' I made some research and I found out this https://github.com/actions/runner-images/issues/6775 Co-authored-by: curquiza <clementine@meilisearch.com>	2023-01-31 12:58:59 +00:00
Tamo	6be9a828fa	makes clippy happy	2023-01-31 13:03:28 +01:00
Tamo	4b7b2d6a90	fix the import of dump v2 generated by meilisearch v0.22.0	2023-01-31 13:03:28 +01:00
bors[bot]	a4e8158239	Merge #774 774: Update version for the next release (v0.41.1) in Cargo.toml files r=curquiza a=meili-bot ⚠️ This PR is automatically generated. Check the new version is the expected one before merging. Co-authored-by: curquiza <curquiza@users.noreply.github.com>	2023-01-31 11:51:42 +00:00
bors[bot]	151e52c481	Merge #3433 3433: Add prototype guide to CONTRIBUTING.md r=curquiza a=curquiza Co-authored-by: curquiza <clementine@meilisearch.com> Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com>	2023-01-31 11:25:46 +00:00
curquiza	e269027cdd	Add git config about ownershio in Docker CI	2023-01-31 12:04:41 +01:00
Loïc Lecrenier	a2690ea8d4	Reduce incremental indexing time of `words_prefix_position_docids` DB This database can easily contain millions of entries. Thus, iterating over it can be very expensive. For regular `documentAdditionOrUpdate` tasks, `del_prefix_fst_words` will always be empty. Thus, we can save a significant amount of time by adding this `if !del_prefix_fst_words.is_empty()` condition. The code's behaviour remains completely unchanged.	2023-01-31 11:42:24 +01:00
bors[bot]	33f61d2cd4	Merge #775 775: Fix clippy for Rust 1.67, allow `uninlined_format_args` r=dureuill a=dureuill # Pull Request milli part of https://github.com/meilisearch/meilisearch/pull/3437 Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-01-31 10:29:24 +00:00
bors[bot]	544b581b15	Merge #3437 3437: Make clippy happy for Rust 1.67, allow uninlined_format_args r=Kerollmops a=dureuill # Pull Request This PR is the equivalent of #3434 for the `release-v1.0.0` branch. See #3434 for more information. Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-01-31 10:29:12 +00:00
f3r10	2922c5c899	Fix code format	2023-01-31 11:28:05 +01:00
f3r10	7681be5367	Format code	2023-01-31 11:28:05 +01:00
f3r10	50bc156257	Fix tests	2023-01-31 11:28:05 +01:00
f3r10	d8207356f4	Skip script,language insertion if language is undetected	2023-01-31 11:28:05 +01:00
f3r10	2d58b28f43	Improve script language codec	2023-01-31 11:28:05 +01:00
f3r10	fd60a39f1c	Format code	2023-01-31 11:28:05 +01:00
f3r10	369c05732e	Add test checking if from script_language_docids database were removed deleted docids	2023-01-31 11:28:05 +01:00
f3r10	34d04f3d3f	Filter from script_language_docids database soft deleted documents	2023-01-31 11:28:05 +01:00
f3r10	a27f329e3a	Add tests for checking that detected script and language associated with document(s) were stored during indexing	2023-01-31 11:28:05 +01:00
f3r10	b216ddba63	Delete and clear data from the new database	2023-01-31 11:28:05 +01:00
f3r10	d97fb6117e	Extract and index data	2023-01-31 11:28:05 +01:00
f3r10	c45d1e3610	Create a new database on index and add a specialized codec for it	2023-01-31 11:28:05 +01:00
Louis Dureuil	5c0668afcf	clippy: allow uninlined_format_args	2023-01-31 11:13:47 +01:00
Louis Dureuil	20f05efb3c	clippy: needless_lifetimes	2023-01-31 11:12:59 +01:00
Louis Dureuil	cbf029f64c	clippy: --fix	2023-01-31 11:12:59 +01:00
curquiza	bffabf9cc6	Update version for the next release (v0.41.1) in Cargo.toml files	2023-01-31 09:56:22 +00:00
bors[bot]	f647b20818	Merge #3434 3434: Make clippy happy for Rust 1.67, allow `uninlined_format_args` r=Kerollmops a=dureuill # Pull Request This PR allows `uninlined_format_args` in CI for clippy. This is due to https://github.com/rust-lang/rust-clippy/issues/10087, which in particular has correctness issues wrt edition 2018 crates, and is a big change altogether. https://github.com/rust-lang/rust-clippy/pull/10265 is already open in order to change the category of this lint to "pedantic", meaning that if this latter PR merges, a future Rust release will accept our code unmodified wrt uninlined format arguments. As a result, this PR introduces the following changes: 1. Allow `uninlined_format_args` in the clippy command in CI. 2. Use rewind rather than seek(0) 3. Remove lifetimes that clippy deems needless. Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-01-31 09:45:12 +00:00
Louis Dureuil	924d5d4c11	clippy: remove needless lifetimes	2023-01-31 10:40:48 +01:00
Louis Dureuil	771a367b97	clippy: use rewind instead of seek 0	2023-01-31 10:40:48 +01:00
Louis Dureuil	07603373f3	clippy: allow uninlined_format_args	2023-01-31 10:15:07 +01:00
Louis Dureuil	d91f8fc493	clippy: Allow uninlined_format_args in CI	2023-01-31 09:56:26 +01:00
Louis Dureuil	3296cf7ae6	clippy: remove needless lifetimes	2023-01-31 09:32:40 +01:00
Louis Dureuil	89675e5f15	clippy: Replace seek 0 by rewind	2023-01-31 09:32:40 +01:00
Louis Dureuil	47b7d515ed	Add more detailed contribution instructions for tests	2023-01-30 17:39:05 +01:00
Clémentine Urquizar - curqui	2ba4629938	Update CONTRIBUTING.md Co-authored-by: Many the fish <many@meilisearch.com>	2023-01-30 15:56:30 +01:00
curquiza	982dd76042	Improve readability	2023-01-30 14:36:22 +01:00
curquiza	3505ee47f8	Add volume to docker command	2023-01-30 14:33:50 +01:00
curquiza	b2d25c07d7	Add guide to create a proto	2023-01-30 14:31:36 +01:00
Clémentine Urquizar - curqui	b9d8bd77fc	Update README.md Co-authored-by: gui machiavelli <hey@guimachiavelli.com>	2023-01-26 18:14:00 +01:00
Clémentine Urquizar - curqui	8a66ba01d8	Update README.md Co-authored-by: gui machiavelli <hey@guimachiavelli.com>	2023-01-26 18:13:53 +01:00
Clémentine Urquizar - curqui	8a6d548041	Update README.md Co-authored-by: gui machiavelli <hey@guimachiavelli.com>	2023-01-26 18:13:08 +01:00
Clémentine Urquizar - curqui	b452358124	Update README.md Co-authored-by: gui machiavelli <hey@guimachiavelli.com>	2023-01-26 18:12:56 +01:00
bors[bot]	bfb1f9279b	Merge #3420 #3422 3420: Add image hyperlink in README.md r=curquiza a=gregsadetsky # Pull Request ## What does this PR do? - tiny README.md improvement: under "SDKs & integration tools", add a hyperlink to the image with all of the language logos so that clicking the image leads to the integrations page. Otherwise, right now, clicking this image leads to the image file in the repo, which is not really useful. ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [X] Have you read the contributing guidelines? - [X] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! 3422: Remove cache from test CI r=dureuill a=curquiza Comment the lines in CIs where we use the test CIs We indeed have cache issues (lack of space on the machine) when running our test CIs https://github.com/meilisearch/meilisearch/pull/3403 Co-authored-by: Greg Sadetsky <lepetitg@gmail.com> Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com> Co-authored-by: curquiza <clementine@meilisearch.com>	2023-01-25 16:25:53 +00:00
Clémentine Urquizar - curqui	48dabd27ea	Update .github/workflows/rust.yml Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-01-25 16:58:15 +01:00
bors[bot]	4549e0a36e	Merge #3415 3415: Test all the errors of wrong `_geo` field and bump milli r=dureuill a=irevoire ## Attention to reviewer The first commit is only a refactoring of the test suite to use snapshot tests everywhere instead of `assert_eq`. It doesn’t change the content of anything and there is probably nothing to review. I just made it for maintenance purpose in the future. Fix https://github.com/meilisearch/meilisearch/issues/3414 Co-authored-by: Tamo <tamo@meilisearch.com>	2023-01-25 15:54:42 +00:00
Tamo	cac93f149e	fix the tests after rebasing	2023-01-25 16:52:54 +01:00
Tamo	481df7a8b6	Update meilisearch/tests/documents/add_documents.rs Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-01-25 16:45:11 +01:00
Tamo	8356f109c1	bump milli to fix the last test	2023-01-25 16:45:11 +01:00
Tamo	934f2b3cb5	exhaustively test all the errors that can arise from a bad geo field	2023-01-25 16:45:11 +01:00
Tamo	a3f1b8fdb9	refactorize the test suite of the add_documents module to use snapshot tests when possible	2023-01-25 16:45:11 +01:00
curquiza	9c3830a19c	Remove cache everywhere	2023-01-25 16:35:02 +01:00
Clémentine Urquizar - curqui	ff6b8dfac4	Remove cache from Windows and macOs CIs	2023-01-25 16:24:04 +01:00
Kerollmops	ec7de4bae7	Make it work for any all routes including stats and index swaps	2023-01-25 16:12:40 +01:00
bors[bot]	d963c2ce55	Merge #3419 3419: Test all the api key error codes r=dureuill a=irevoire Partially fix #3325 Co-authored-by: Tamo <tamo@meilisearch.com>	2023-01-25 15:09:19 +00:00
bors[bot]	5beb1aab7d	Merge #3418 3418: Compute the size of the auth-controller, index-scheduler and all update files in the global stats r=dureuill a=irevoire Fix https://github.com/meilisearch/meilisearch/issues/3201 Co-authored-by: Tamo <tamo@meilisearch.com>	2023-01-25 14:05:17 +00:00
Kerollmops	184b8afd9e	Make it work in the CreateApiKey struct	2023-01-25 15:01:50 +01:00
Tamo	a858531574	apply review comments	2023-01-25 14:51:36 +01:00
Kerollmops	29961b8c6b	Make it work with the dumps	2023-01-25 14:41:36 +01:00
Clément Renault	0b08413c98	Introduce the IndexUidPattern type	2023-01-25 14:22:17 +01:00
Clément Renault	474d4ec498	Add tests for the index patterns	2023-01-25 14:22:16 +01:00
Tamo	bf94f89035	Update index-scheduler/src/lib.rs Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-01-25 11:31:50 +01:00
Tamo	3bcff60d1c	makes clippy happy	2023-01-25 11:31:48 +01:00
Tamo	04c4487660	udpate the analytics with the new stats method	2023-01-25 11:25:04 +01:00
Tamo	c92948b143	Compute the size of the auth-controller, index-scheduler and all update files in the global stats	2023-01-25 11:25:02 +01:00
bors[bot]	0544b60974	Merge #3409 3409: Bump libgit2-sys from 0.14.1+1.5.0 to 0.14.2+1.5.1 r=Kerollmops a=dependabot[bot] Bumps [libgit2-sys](https://github.com/rust-lang/git2-rs) from 0.14.1+1.5.0 to 0.14.2+1.5.1. <details> <summary>Commits</summary> <ul> <li><a href="`a233483a39`"><code>a233483</code></a> Update to libgit2 1.5.1</li> <li><a href="`bce15556ef`"><code>bce1555</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/rust-lang/git2-rs/issues/909">#909</a> from ehuss/ssh-keys</li> <li><a href="`222fbf3b9e`"><code>222fbf3</code></a> Bump versions</li> <li><a href="`fa41943135`"><code>fa41943</code></a> Change the certificate_check callback to support passthrough.</li> <li><a href="`84e21aad4e`"><code>84e21aa</code></a> Add ability to get the SSH host key and its type.</li> <li><a href="`e6aa6666b9`"><code>e6aa666</code></a> Bump git2-curl version. (<a href="https://github-redirect.dependabot.com/rust-lang/git2-rs/issues/861">#861</a>)</li> <li><a href="`46674cebd9`"><code>46674ce</code></a> Fix warning about unused_must_use for Box::from_raw (<a href="https://github-redirect.dependabot.com/rust-lang/git2-rs/issues/860">#860</a>)</li> <li><a href="`951dce9dea`"><code>951dce9</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/rust-lang/git2-rs/issues/858">#858</a> from davidkna/git2150</li> <li><a href="`8871f8e9b3`"><code>8871f8e</code></a> bump libgit2 to 1.5.0</li> <li><a href="`04278a24ba`"><code>04278a2</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/rust-lang/git2-rs/issues/839">#839</a> from davidkna/libgit2_143</li> <li>Additional commits viewable in <a href="https://github.com/rust-lang/git2-rs/compare/0.14.1...libgit2-sys-0.14.2">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=libgit2-sys&package-manager=cargo&previous-version=0.14.1+1.5.0&new-version=0.14.2+1.5.1)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) Dependabot will resolve any conflicts with this PR as long as you don't alter it yourself. You can also trigger a rebase manually by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) You can disable automated security fix PRs for this repo from the [Security Alerts page](https://github.com/meilisearch/meilisearch/network/alerts). </details> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2023-01-25 09:34:34 +00:00
Greg Sadetsky	4223c51838	Add image hyperlink in README.md	2023-01-24 15:24:09 -05:00
bors[bot]	b3c2a4ae27	Merge #3412 3412: When adding documents, trying to update the primary-key now throw an error r=Kerollmops a=irevoire While updating the test suite, I also noticed an issue with the indexed_documents value of failed tasks and had to update it. I also named a bunch of snapshots that had no name, sorry 😬 Fixes https://github.com/meilisearch/meilisearch/issues/3385 Fixes https://github.com/meilisearch/meilisearch/issues/3411 Co-authored-by: Tamo <tamo@meilisearch.com>	2023-01-24 17:14:11 +00:00
Tamo	c7b2e3be87	apply review comments	2023-01-24 17:54:43 +01:00
Tamo	aa17a54feb	test all the api key error codes	2023-01-24 17:30:35 +01:00
bors[bot]	6f71a2b38b	Merge #3403 3403: Add `--all` to test CI r=curquiza a=curquiza Discussed with `@irevoire` [here](https://meilisearch.slack.com/archives/G01A1F4KVGU/p1674144546920649?thread_ts=1674144456.561199&cid=G01A1F4KVGU) (internal link) Co-authored-by: curquiza <clementine@meilisearch.com>	2023-01-24 16:23:08 +00:00
bors[bot]	898160587f	Merge #3416 3416: Add tests on the index resource r=Kerollmops a=irevoire Fix part of https://github.com/meilisearch/meilisearch/issues/3325 Co-authored-by: Tamo <tamo@meilisearch.com>	2023-01-24 15:26:18 +00:00
bors[bot]	7c9935f96a	Merge #769 769: Modify README to prevent contributions r=Kerollmops a=curquiza Co-authored-by: curquiza <clementine@meilisearch.com> Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com>	2023-01-24 15:14:31 +00:00
Clémentine Urquizar - curqui	f7ae8bc065	Update README.md Co-authored-by: Clément Renault <clement@meilisearch.com>	2023-01-24 15:58:41 +01:00
Clémentine Urquizar - curqui	3d8a3d22d1	Update README.md Co-authored-by: Clément Renault <clement@meilisearch.com>	2023-01-24 15:58:34 +01:00
bors[bot]	30f88350c7	Merge #773 773: bump milli r=Kerollmops a=irevoire I need a new release of milli for https://github.com/meilisearch/meilisearch/pull/3415 Co-authored-by: Tamo <tamo@meilisearch.com>	2023-01-24 14:51:32 +00:00
bors[bot]	4c4baaf1ce	Merge #3387 3387: Update create-issue-dependencies.yml r=Kerollmops a=curquiza Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com>	2023-01-24 14:44:09 +00:00
Tamo	55e8046551	bump milli	2023-01-24 13:52:21 +01:00
Tamo	32364e9919	add tests on the index resource	2023-01-24 13:20:20 +01:00
bors[bot]	4e4d8dfda7	Merge #772 772: Throw an error on unknown fields specified in the _geo field r=irevoire a=irevoire Fix parts of https://github.com/meilisearch/meilisearch/issues/3414 Co-authored-by: Tamo <tamo@meilisearch.com>	2023-01-24 11:36:00 +00:00
Tamo	de3c4f1986	throw an error on unknown fields specified in the _geo field	2023-01-24 12:23:24 +01:00
Tamo	ea3b269b77	reformat	2023-01-23 23:59:34 +01:00
Tamo	a4be4c49e8	Update index-scheduler/src/batch.rs Co-authored-by: Clément Renault <clement@meilisearch.com>	2023-01-23 23:58:03 +01:00
Tamo	7d1ebb7295	add test on the autobatcher layer	2023-01-23 20:56:12 +01:00
bors[bot]	e664f09045	Merge #3396 3396: Update our error message about negative integer r=dureuill a=irevoire Fix https://github.com/meilisearch/meilisearch/issues/3394 Co-authored-by: Tamo <tamo@meilisearch.com>	2023-01-23 19:50:04 +00:00
Tamo	767cb725a5	reimplement the batching of task with or without primary key in the autobatcher	2023-01-23 20:18:22 +01:00
Tamo	13c2cd700d	Update error message about negative integer	2023-01-23 18:08:46 +01:00
bors[bot]	fea41ca788	Merge #3404 3404: Fix matching strategy error r=irevoire a=ManyTheFish # Pull Request ## Related issue Fixes #3391 Co-authored-by: ManyTheFish <many@meilisearch.com> Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-01-23 17:04:05 +00:00
bors[bot]	217504fff3	Merge #3406 3406: Master Key: Implements errors and warnings from the specification r=irevoire a=dureuill <sub>Now in technicolor</sub> # Pull Request ## What does this PR do? - Uses `atty` and `termcolor` as dependency - Use these dependencies to print colored background for warning messages - Update messages to match https://github.com/meilisearch/specifications/pull/209 ## PR checklist Please check if your PR fulfills the following requirements: - [ ] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [ ] Have you read the contributing guidelines? - [ ] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-01-23 16:39:18 +00:00
Tamo	5672118bfa	When adding documents, trying to update the primary-key now throw an error While updating the test suite I also noticed an issue with the indexed_documents value of failed task and had to update it. I also named a bunch of snapshots that had no name sorry 😬	2023-01-23 17:32:13 +01:00
Louis Dureuil	57682cbabe	Fix test url after #3398	2023-01-23 15:43:17 +01:00
ManyTheFish	5dd582918d	Add test	2023-01-23 15:40:42 +01:00
bors[bot]	74747b65b1	Merge #3395 3395: Indicate filterable attributes in facet distributions when user requests a non filterable one. r=irevoire a=dureuill # Pull Request ## Related issue Fixes #3390 ## What does this PR do? - bump milli & deserr - Update and add tests ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Louis Dureuil <louis@meilisearch.com> Co-authored-by: ManyTheFish <many@meilisearch.com> Co-authored-by: Tamo <tamo@meilisearch.com>	2023-01-23 13:53:55 +00:00
Tamo	c79b6a1ee4	bump milli	2023-01-23 14:13:19 +01:00
ManyTheFish	f0e6b9c0c5	Update deserr to 0.3.0	2023-01-23 14:13:04 +01:00
Louis Dureuil	56db54486c	Add tests	2023-01-23 14:00:30 +01:00
Louis Dureuil	a9b3f91467	Add missing space Co-authored-by: Guillaume Mourier <guillaume@meilisearch.com>	2023-01-23 10:33:30 +01:00
dependabot[bot]	5f4497935f	Bump libgit2-sys from 0.14.1+1.5.0 to 0.14.2+1.5.1 Bumps [libgit2-sys](https://github.com/rust-lang/git2-rs) from 0.14.1+1.5.0 to 0.14.2+1.5.1. - [Release notes](https://github.com/rust-lang/git2-rs/releases) - [Commits](https://github.com/rust-lang/git2-rs/compare/0.14.1...libgit2-sys-0.14.2) --- updated-dependencies: - dependency-name: libgit2-sys dependency-type: indirect ... Signed-off-by: dependabot[bot] <support@github.com>	2023-01-20 23:38:40 +00:00
Gregory Conrad	3f69dd6450	feat: add Cargo feature for LMDB's POSIX semaphores	2023-01-19 12:08:38 -05:00
bors[bot]	1c4b1b3b2d	Merge #770 770: Update deserr v0.3.0 r=irevoire a=ManyTheFish related to https://github.com/meilisearch/meilisearch/issues/3391 Co-authored-by: Many the fish <many@meilisearch.com>	2023-01-19 17:05:56 +00:00
Louis Dureuil	0de9a3ffe7	Implements errors and warnings from the specification Now in technicolor	2023-01-19 18:04:45 +01:00
bors[bot]	b4f1e9bc36	Merge #771 771: Update version for the next release (v0.40.0) in Cargo.toml files r=curquiza a=meili-bot ⚠️ This PR is automatically generated. Check the new version is the expected one before merging. Co-authored-by: curquiza <curquiza@users.noreply.github.com>	2023-01-19 16:45:20 +00:00
curquiza	abd65d9307	Update version for the next release (v0.40.0) in Cargo.toml files	2023-01-19 16:43:45 +00:00
Many the fish	30fc376713	Update deserr v0.3.0	2023-01-19 17:37:30 +01:00
curquiza	2a1787ed22	Add --all in test CI	2023-01-19 17:26:47 +01:00
curquiza	d1a31afdd6	Modify README to prevent contributions	2023-01-19 17:17:34 +01:00
bors[bot]	60018d0fe4	Merge #3343 3343: Extract creation and last updated timestamp for v3 dump r=curquiza a=FrancisMurillo # Pull Request ## Related issue Fixes #2988 ## What does this PR do? Inspired by the v4 dump implementation, this extracts the first `createdAt` and last `updatedAt` fields by parsing the task queue. Questions: - Should the parsing of the tasks be cached instead of being parsed for every index since it might add a performance penalty? - I am not sure if the `created_at` and `processed_at` fields are correct - Should I assume the data is sorted in some order like with `uuid` or `updateId`? I assumed the list is unordered. - I was planning to populate my dev instance with data and dump my data. Is there a way to dump with previous versions? ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Francis Murillo <evacuee.overlap.vs3op@aleeas.com>	2023-01-19 16:14:21 +00:00
bors[bot]	8fb685f5aa	Merge #3401 3401: improve the error messages for the immutable fields r=dureuill a=irevoire Fix https://github.com/meilisearch/meilisearch/issues/3400 Co-authored-by: Tamo <tamo@meilisearch.com>	2023-01-19 15:52:50 +00:00
Tamo	e3742a38d4	improve the error messages for the immutable fields	2023-01-19 16:49:44 +01:00
curquiza	13b1abceaf	Rework technical information in the README	2023-01-19 16:23:06 +01:00
bors[bot]	e16b5c615a	Merge #3398 3398: Error links use underscores again r=irevoire a=dureuill # Pull Request ## Related issue Follow-up of #3288 where [it was decided](https://github.com/meilisearch/meilisearch/pull/3288#issuecomment-1396733603) to revert course on the separator to use in error anchors. ## What does this PR do? - Use `_` again as separator in anchors of error link - Fix tests Impacts `@meilisearch/docs-team` : we need `_`-separated anchors to be generated in the online documentation to match the ones emitted from the engine. Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-01-19 15:17:33 +00:00
bors[bot]	3521a3a0b2	Merge #763 763: Fixes error message when lat and lng are unparseable r=loiclec a=ahlner # Pull Request ## Related issue Fixes partially [#3007](https://github.com/meilisearch/meilisearch/issues/3007) ## What does this PR do? - Changes function validate_geo_from_json to return a BadLatitudeAndLongitude if lat or lng is a string and not parseable to f64 - implemented some unittests - Derived PartialEq for GeoError to use assert_eq! in tests ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Philipp Ahlner <philipp@ahlner.com>	2023-01-19 15:15:46 +00:00
Louis Dureuil	d2420f5c8f	Fix non insta tests	2023-01-19 16:10:05 +01:00
Louis Dureuil	72e2b220ed	Fix tests	2023-01-19 15:48:20 +01:00
bors[bot]	40a53f8824	Merge #767 767: Update version for the next release (v0.39.2) in Cargo.toml files r=curquiza a=meili-bot ⚠️ This PR is automatically generated. Check the new version is the expected one before merging. Co-authored-by: curquiza <curquiza@users.noreply.github.com>	2023-01-19 14:48:12 +00:00
Louis Dureuil	b0c33ed6d2	Error codes are underscore again	2023-01-19 15:47:01 +01:00
Philipp Ahlner	f5ca421227	Superfluous test removed	2023-01-19 15:39:21 +01:00
curquiza	3f048927a0	Update version for the next release (v0.39.2) in Cargo.toml files	2023-01-19 14:29:09 +00:00
bors[bot]	e7c0617699	Merge #766 766: Indicate filterable attributes when the user sets a non filterable attribute in facet distributions r=irevoire a=dureuill # Pull Request ## Related issue Related to https://github.com/meilisearch/meilisearch/issues/3390 ## What does this PR do? - Title ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-01-19 14:18:13 +00:00
bors[bot]	a1e9c44fe5	Merge #3389 3389: Return `invalid_search_facets` rather than `bad_request` when using facet on a non filterable attribute r=irevoire a=dureuill # Pull Request ## Related issue Fixes https://github.com/meilisearch/meilisearch/issues/3384 ## What does the PR does - title - also adds a test ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-01-19 13:19:22 +00:00
bors[bot]	7df1dda002	Merge #3393 3393: improve the error message when no task filter are specified for the cancelation or deletion of tasks r=dureuill a=irevoire Close https://github.com/meilisearch/meilisearch/issues/3392 Was already present in v0.30 Co-authored-by: Tamo <tamo@meilisearch.com>	2023-01-19 12:55:52 +00:00
Louis Dureuil	3d8ca62c35	InvalidFacetDistribution returns invalid_search_facet	2023-01-19 13:41:26 +01:00
Tamo	e8e7070cc6	improve the error message when no task filter are specified for the cancelation or deletion of tasks	2023-01-19 12:42:08 +01:00
Francis Murillo	798aa4ee92	Fix clippy issues	2023-01-19 19:38:20 +08:00
Louis Dureuil	4fd6fd9bef	Indicate filterable attributes when the user set a non filterable attribute in facet distributions	2023-01-19 12:25:18 +01:00
bors[bot]	f857d9c2df	Merge #3383 3383: Fix api key patch r=irevoire a=irevoire This was introduced in the previous rc Fix https://github.com/meilisearch/meilisearch/issues/3374 Co-authored-by: Tamo <tamo@meilisearch.com>	2023-01-19 10:05:09 +00:00
Philipp Ahlner	a2cd7214f0	Fixes error message when lat/lng are unparseable	2023-01-19 10:10:26 +01:00
Clémentine Urquizar - curqui	0ce1d6d487	Update create-issue-dependencies.yml	2023-01-18 23:43:33 +01:00
Tamo	d0988e115f	fix the patch of description and name for the api-key	2023-01-18 19:07:26 +01:00
Tamo	5dcb920fb4	improve the tests	2023-01-18 18:27:00 +01:00
bors[bot]	b3166df7ea	Merge #3372 3372: Enhance facet string normalization r=ManyTheFish a=ManyTheFish # Pull Request Use compatibility decomposition normalizer in facet string extraction in order to have a more human friendly sort order. Now, [é (U+00E9)](https://www.compart.com/fr/unicode/U+00E9) is converted to [e (U+0065)](https://www.compart.com/fr/unicode/U+0065) + [◌́ (U+0301)](https://www.compart.com/fr/unicode/U+0301). This way any word starting with an accented/diacritized version of a character is put just after the words starting with the unaccented version of the character. ## Related issue Fixes #3260 Co-authored-by: ManyTheFish <many@meilisearch.com>	2023-01-18 17:17:53 +00:00
bors[bot]	6f7e0c431a	Merge #3341 3341: add functionnal + error tests on the swap_indexes route and fix a confusing error message r=loiclec a=irevoire Fix https://github.com/meilisearch/meilisearch/issues/3340 Fix part of https://github.com/meilisearch/meilisearch/issues/3325 Fix https://github.com/meilisearch/meilisearch/issues/3381 Test both the functionality and the error codes Co-authored-by: Tamo <tamo@meilisearch.com> Co-authored-by: Loïc Lecrenier <loic.lecrenier@me.com>	2023-01-18 16:32:22 +00:00
Tamo	00f6af6475	fix a wrong error message	2023-01-18 17:26:48 +01:00
bors[bot]	1803998017	Merge #3335 #3353 3335: Remove test badge r=curquiza a=curquiza Suggestion of removal: from my point of view, this badge does not provide any useful information, and most of all is often outdated. Currently ours is "no status" despite our tests passing. Plus, sometimes our tests are not passing because we are still in development, but it does not mean our current provided binaries are not. <img width="619" alt="Capture d’écran 2023-01-12 à 14 06 40" src="https://user-images.githubusercontent.com/20380692/212074200-f9e3ab3e-ad1d-4171-bd13-46584c3cd117.png"> 3353: Bump svenstaro/upload-release-action from 2.3.0 to 2.4.0 r=curquiza a=dependabot[bot] Bumps [svenstaro/upload-release-action](https://github.com/svenstaro/upload-release-action) from 2.3.0 to 2.4.0. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/svenstaro/upload-release-action/releases">svenstaro/upload-release-action's releases</a>.</em></p> <blockquote> <h2>2.4.0</h2> <ul> <li>Update to node 16</li> <li>Bump most dependencies</li> </ul> </blockquote> </details> <details> <summary>Changelog</summary> <p><em>Sourced from <a href="https://github.com/svenstaro/upload-release-action/blob/master/CHANGELOG.md">svenstaro/upload-release-action's changelog</a>.</em></p> <blockquote> <h2>[2.4.0] - 2023-01-09</h2> <ul> <li>Update to node 16</li> <li>Bump most dependencies</li> </ul> </blockquote> </details> <details> <summary>Commits</summary> <ul> <li><a href="`2728235f7d`"><code>2728235</code></a> 2.4.0</li> <li><a href="`c2e0608dc4`"><code>c2e0608</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/svenstaro/upload-release-action/issues/88">#88</a> from svenstaro/dependabot/npm_and_yarn/json5-1.0.2</li> <li><a href="`bd74772a1a`"><code>bd74772</code></a> Don't shadow vars</li> <li><a href="`16e7903b2d`"><code>16e7903</code></a> Bump json5 from 1.0.1 to 1.0.2</li> <li><a href="`f2c549b117`"><code>f2c549b</code></a> Bump some more deps</li> <li><a href="`7a7d004438`"><code>7a7d004</code></a> Bump some deps</li> <li><a href="`9c4a92ec0d`"><code>9c4a92e</code></a> Use explicit any</li> <li><a href="`039214a996`"><code>039214a</code></a> Bump jest and typescript versions</li> <li><a href="`2b373356cb`"><code>2b37335</code></a> Update to node16</li> <li><a href="`fb1eb39e74`"><code>fb1eb39</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/svenstaro/upload-release-action/issues/75">#75</a> from svenstaro/dependabot/npm_and_yarn/jsdom-16.7.0</li> <li>Additional commits viewable in <a href="https://github.com/svenstaro/upload-release-action/compare/2.3.0...2.4.0">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=svenstaro/upload-release-action&package-manager=github_actions&previous-version=2.3.0&new-version=2.4.0)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) You can trigger a rebase of this PR by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) </details> Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2023-01-18 15:45:21 +00:00
bors[bot]	3e5b3df487	Merge #3370 #3373 #3375 3370: make the swap indexes not found errors return an IndexNotFound error-code r=irevoire a=irevoire Fix https://github.com/meilisearch/meilisearch/issues/3368 3373: fix a wrong error code and add tests on the document resource r=irevoire a=irevoire Fix https://github.com/meilisearch/meilisearch/issues/3371 3375: Avoid deleting all task invalid canceled by r=irevoire a=Kerollmops Fixes #3369 by making sure that at least one `canceledBy` task filter parameter matches something. Co-authored-by: Tamo <tamo@meilisearch.com> Co-authored-by: Kerollmops <clement@meilisearch.com>	2023-01-18 15:21:11 +00:00
Kerollmops	e89973f1bf	Do not delete all tasks when no canceled-by matches	2023-01-18 15:50:46 +01:00
Kerollmops	d3c796af38	Add a new test to check that invalid canceledBy works correctly	2023-01-18 15:50:46 +01:00
Kerollmops	182eea1f17	Introduce a canceledBy filter for the tests	2023-01-18 15:50:42 +01:00
bors[bot]	1af3089456	Merge #3348 3348: fix cargo flaky r=irevoire a=irevoire Partially fix https://github.com/meilisearch/meilisearch/issues/3273 Ideally, we should revert this commit and fix cargo-flaky directly to ensure we never forget to add a sub-crate to the CI. ----- Here is an example of the CI running (and thus working); https://github.com/meilisearch/meilisearch/actions/runs/3932783699/jobs/6725755801 Co-authored-by: Tamo <tamo@meilisearch.com>	2023-01-18 14:46:50 +00:00
Tamo	a4476c20f8	fix a wrong error code and add tests on the document resource	2023-01-18 15:28:02 +01:00
ManyTheFish	d1fc42b53a	Use compatibility decomposition normalizer in facets	2023-01-18 15:02:13 +01:00
ManyTheFish	e64571a881	Add test sorting string with diacritics	2023-01-18 14:43:38 +01:00
Tamo	57da80900d	make the swap indexes not found errors return an IndexNotFound error code	2023-01-18 14:16:00 +01:00
bors[bot]	7322f4e78e	Merge #3355 3355: fix the wrong error code on minWordSizeForTypos r=irevoire a=irevoire Fix #3354 Co-authored-by: Tamo <tamo@meilisearch.com>	2023-01-18 12:25:03 +00:00
Philipp Ahlner	497187083b	Add test for bug #3007 : Wrong error message Adds a test for #3007: Wrong error message when lat and lng are unparseable	2023-01-18 13:24:26 +01:00
Tamo	0f727d079b	fix the wrong error code on minWordSizeForTypos	2023-01-18 12:28:46 +01:00
dependabot[bot]	32e2848a74	Bump svenstaro/upload-release-action from 2.3.0 to 2.4.0 Bumps [svenstaro/upload-release-action](https://github.com/svenstaro/upload-release-action) from 2.3.0 to 2.4.0. - [Release notes](https://github.com/svenstaro/upload-release-action/releases) - [Changelog](https://github.com/svenstaro/upload-release-action/blob/master/CHANGELOG.md) - [Commits](https://github.com/svenstaro/upload-release-action/compare/2.3.0...2.4.0) --- updated-dependencies: - dependency-name: svenstaro/upload-release-action dependency-type: direct:production update-type: version-update:semver-minor ... Signed-off-by: dependabot[bot] <support@github.com>	2023-01-18 11:02:44 +00:00
bors[bot]	6b3da8a6de	Merge #3346 3346: Import milli 🎉 r=Kerollmops a=Kerollmops Fixes https://github.com/meilisearch/meilisearch/issues/2901 Main work - integrate the milli repository as an internal crate into this repo - Update the Cargo.toml accordingly - Ensure meilisearch-type now uses the internal milli crate and not the remote repository - Update the milli's version to follow the meilisearch one Also - Removed the beta tests in test CI (will be re-integrated later if needed) - Move and modify milli's README into the `milli` folder - remove the script folder from `milli` - Removed useless CI (release-drafter and enforce-label) ⚠️ Also, import all the `release-v1.0.0` until [a5c4fb](`a5c4fbbcea`) included (merged of the PR https://github.com/meilisearch/meilisearch/pull/3334) Co-authored-by: Loïc Lecrenier <loic.lecrenier@me.com> Co-authored-by: bors[bot] <26634292+bors[bot]@users.noreply.github.com> Co-authored-by: Samyak S Sarnayak <samyak201@gmail.com> Co-authored-by: unvalley <kirohi.code@gmail.com> Co-authored-by: Samyak Sarnayak <samyak201@gmail.com>	2023-01-18 10:19:42 +00:00
Clément Renault	0769090dd6	Add a note in the README about the crates versionning	2023-01-18 10:08:12 +01:00
Loïc Lecrenier	82bdb54537	Update the index swap tests after git rebase	2023-01-18 09:40:41 +01:00
Tamo	b6ec1f1c6d	add functionnal + error tests on the swap_indexes route	2023-01-18 09:36:04 +01:00
Clément Renault	1d507c84b2	Fix the formatting	2023-01-17 18:25:55 +01:00
Clément Renault	1b78231e18	Make clippy happy	2023-01-17 18:25:54 +01:00
Clément Renault	2b1f6a7f11	Fix the CI to ignore a missing file	2023-01-17 16:26:03 +01:00
Francis Murillo	6993924f32	Use finished_at for v3 dumps instead	2023-01-17 23:11:49 +08:00
bors[bot]	41a970247e	Merge #3339 3339: Continued deserr integration r=irevoire a=loiclec Fix https://github.com/meilisearch/meilisearch/issues/3337 Fix https://github.com/meilisearch/meilisearch/issues/3338 1. Add new error codes that should have been implemented earlier: - `MissingApiKeyActions` - `MissingApiKeyExpiresAt` - `MissingApiKeyIndexes` - `MissingSwapIndexes` 2. Fix a bug where it was possible to create an API key without specifying the value of `expiresAt` 3. Improve the error messages generated by deserr. Have specific error messages for JSON and QueryParam deserialisation errors. 4. Improve error tests by passing query params as arguments to `GET` routes directly instead of using an intermediary JSON object 5. [Use invalid_index_uid error code in more places](`e225608337`) Co-authored-by: Loïc Lecrenier <loic.lecrenier@me.com>	2023-01-17 14:41:22 +00:00
Loïc Lecrenier	e225608337	Use invalid_index_uid error code in more places	2023-01-17 15:28:06 +01:00
Loïc Lecrenier	56e79fa850	Update task snapshot test and clean up details	2023-01-17 13:19:04 +01:00
Loïc Lecrenier	c71a8ea183	Update to latest milli and deserr	2023-01-17 13:10:38 +01:00
bors[bot]	0c7d1f761e	Merge #765 765: Update version for the next release (v0.39.1) in Cargo.toml files r=curquiza a=meili-bot ⚠️ This PR is automatically generated. Check the new version is the expected one before merging. Co-authored-by: curquiza <curquiza@users.noreply.github.com>	2023-01-17 11:04:26 +00:00
curquiza	e3d30e28ef	Update version for the next release (v0.39.1) in Cargo.toml files	2023-01-17 10:50:29 +00:00
bors[bot]	63af1e9f28	Merge #764 764: Update deserr to latest version r=irevoire a=loiclec Update deserr to 0.1.5, which changes the `DeserializeFromValue` trait, getting rid of the `default()` method. Co-authored-by: Loïc Lecrenier <loic.lecrenier@me.com>	2023-01-17 10:39:36 +00:00
Loïc Lecrenier	f073a86387	Update deserr to latest version	2023-01-17 11:28:19 +01:00
Loïc Lecrenier	b781f9a0f9	cargo fmt	2023-01-17 11:07:07 +01:00
Loïc Lecrenier	07b90dec08	Remove unused proptest dependency	2023-01-17 11:07:07 +01:00
Loïc Lecrenier	9194508a0f	Refactor query parameter deserialisation logic	2023-01-17 11:07:07 +01:00
Tamo	9dd01ff44b	fix cargo flaky	2023-01-17 11:03:44 +01:00
Loïc Lecrenier	49ddaaef49	Fix missing_swap_indexes error code and handling of expires_at param... of create api key route	2023-01-17 09:43:07 +01:00
Loïc Lecrenier	766dd830ae	Update deserr to latest version + add new error codes for missing fields - missing_api_key_indexes - missing_api_key_actions - missing_api_key_expires_at - missing_swap_indexes_indexes	2023-01-17 09:43:07 +01:00
Loïc Lecrenier	436ae4e466	Improve error messages generated by deserr Split Json and Query Parameter error types	2023-01-17 09:43:07 +01:00
Kerollmops	507a7bad96	Use the local milli subcrate	2023-01-16 17:35:54 +01:00
Kerollmops	cde62fcb5b	Merge remote-tracking branch 'origin/release-v1.0.0' into import-milli	2023-01-16 17:35:18 +01:00
Kerollmops	03a82136dc	Remove the useless cli subcrate	2023-01-16 17:08:43 +01:00
Kerollmops	e68758cec4	Refine the cargo workspace profile settings	2023-01-16 17:04:25 +01:00
Kerollmops	4fb47492e5	Make clippy happy	2023-01-16 16:35:58 +01:00
Kerollmops	5bab8cf7ec	Remove useless CI configs	2023-01-16 16:31:46 +01:00
Kerollmops	97005dd505	Bump the milli-imported crates to v1.0.0	2023-01-16 16:29:12 +01:00
Kerollmops	eabef5194a	Remove the useless script folder	2023-01-16 16:26:07 +01:00
Kerollmops	ebb2494879	Add a README to the milli crate	2023-01-16 16:25:12 +01:00
Kerollmops	0cec352d2b	Merge remote-tracking branch 'milli/main' into import-milli	2023-01-16 16:20:22 +01:00
Francis Murillo	a97281af08	Extract createdAt and updatedAt from v3 dump	2023-01-13 22:45:45 +08:00
bors[bot]	a5c4fbbcea	Merge #3334 3334: Add specific error codes `immutable_...` r=irevoire a=loiclec Add the following error codes: When an immutable field of API key is sent to the `PATCH /keys` route: - `ImmutableApiKeyUid` - `ImmutableApiKeyKey` - `ImmutableApiKeyActions` - `ImmutableApiKeyIndexes` - `ImmutableApiKeyExpiresAt` - `ImmutableApiKeyCreatedAt` - `ImmutableApiKeyUpdatedAt` When an immutable field of Index is sent to the `PATCH /indexes/{uid}` route: - `ImmutableIndexUid` - `ImmutableIndexCreatedAt` - `ImmutableIndexUpdatedAt` Co-authored-by: Loïc Lecrenier <loic.lecrenier@me.com> Co-authored-by: Tamo <tamo@meilisearch.com>	2023-01-12 15:31:38 +00:00
Tamo	21b8cd53b7	reformat	2023-01-12 16:20:24 +01:00
Loïc Lecrenier	7f80b116bc	Add specific immutable_field error codes	2023-01-12 16:20:14 +01:00
bors[bot]	341f8478b4	Merge #3330 3330: test the error codes on the task routes + fix the missing error codes on the limit and from r=dureuill a=irevoire Co-authored-by: Tamo <tamo@meilisearch.com> Co-authored-by: Loïc Lecrenier <loic.lecrenier@me.com>	2023-01-12 15:02:44 +00:00
Tamo	79c7f65c30	make a test more reliable	2023-01-12 15:39:28 +01:00
bors[bot]	2bc60c29fc	Merge #3336 3336: Add missing `needs:` to the git latest tag workflow r=curquiza a=curquiza Fixes this problem: the workflow to update the latest git tag was triggered despite the first check failed <img width="580" alt="Capture d’écran 2023-01-12 à 15 07 00" src="https://user-images.githubusercontent.com/20380692/212087926-975eb387-c8c9-4789-8a62-a56143b9bbd4.png"> These leads to update our latest git tag: our latest git tag corresponds to the `v1.0.0-rc.0` tag instead of `v0.30.5`. (I'm fixing this right now) <img width="586" alt="Capture d’écran 2023-01-12 à 15 08 15" src="https://user-images.githubusercontent.com/20380692/212088136-f4bc2e9c-d824-4c23-8213-52598c742ebd.png"> Co-authored-by: curquiza <clementine@meilisearch.com>	2023-01-12 14:24:31 +00:00
curquiza	680ea39bba	Add missingneeds: to the git latest tag workflow	2023-01-12 15:04:11 +01:00
Tamo	a524dfb713	fix the analytics	2023-01-12 14:49:50 +01:00
Tamo	705fcaa3b8	reformat the imports	2023-01-12 14:09:15 +01:00
Clémentine Urquizar - curqui	55605435bc	Remove test badge	2023-01-12 14:04:48 +01:00
Loïc Lecrenier	a09b6a341d	Move tasks route to deserr	2023-01-12 13:57:29 +01:00
Tamo	387874ea26	test the error codes on the task routes	2023-01-12 13:46:19 +01:00
bors[bot]	5c1a7c3b9a	Merge #3329 3329: Refactor error handling from deserr r=irevoire a=loiclec Close https://github.com/meilisearch/meilisearch/issues/3318 Close https://github.com/meilisearch/meilisearch/issues/3289 [TODO] Co-authored-by: Loïc Lecrenier <loic.lecrenier@me.com> Co-authored-by: Tamo <tamo@meilisearch.com>	2023-01-11 18:15:32 +00:00
Tamo	6d658f4c52	fix a wrong error code + update some error messages	2023-01-11 19:14:11 +01:00
Tamo	bf573885ea	integrate the latest version of milli	2023-01-11 19:08:39 +01:00
Tamo	a68ac3a1dc	reformat the headers	2023-01-11 19:08:39 +01:00
Tamo	b252c87197	add tests on the sub settings routes	2023-01-11 19:08:39 +01:00
Loïc Lecrenier	b0b7ad7caf	Apply review suggestions	2023-01-11 19:08:39 +01:00
Loïc Lecrenier	c91ffec72e	Update Cargo.toml	2023-01-11 19:08:39 +01:00
Loïc Lecrenier	1fc11264e8	Refactor deserr integration	2023-01-11 19:08:39 +01:00
Loïc Lecrenier	2bc2e99ff3	Simplify declaration of the error codes	2023-01-11 19:08:39 +01:00
bors[bot]	808e184069	Merge #3324 3324: Add a test on the search route for each possible error codes r=irevoire a=irevoire Co-authored-by: Tamo <tamo@meilisearch.com>	2023-01-11 16:08:19 +00:00
bors[bot]	e6bea99974	Merge #762 762: Update version for the next release (v0.39.0) in Cargo.toml files r=curquiza a=meili-bot ⚠️ This PR is automatically generated. Check the new version is the expected one before merging. Co-authored-by: curquiza <curquiza@users.noreply.github.com>	2023-01-11 15:07:33 +00:00
curquiza	9e32ac7cb2	Update version for the next release (v0.39.0) in Cargo.toml files	2023-01-11 15:05:06 +00:00
bors[bot]	302d6cccd7	Merge #761 761: Integrate deserr r=irevoire a=loiclec 1. `Setting<T>` now implements `DeserializeFromValue` 2. The settings now store ranking rules as strongly typed `Criterion` instead of `String`, since the validation of the ranking rules will be done on meilisearch's side from now on Co-authored-by: Loïc Lecrenier <loic.lecrenier@me.com>	2023-01-11 14:35:15 +00:00
bors[bot]	21b7d709ad	Merge #759 759: Change primary key inference error messages r=Kerollmops a=dureuill # Pull Request ## Related issue Milli part of https://github.com/meilisearch/meilisearch/issues/3301 ## What does this PR do? - Change error message strings ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-01-11 14:04:25 +00:00
Tamo	7a30d98264	fix a flaky test	2023-01-11 14:54:29 +01:00
Loïc Lecrenier	02fd06ea0b	Integrate deserr	2023-01-11 13:56:47 +01:00
Tamo	d0a85057a3	fix the bad filter test	2023-01-11 11:37:12 +01:00
bors[bot]	b3574de809	Merge #3321 3321: Update the system http error code to return an internal server error r=irevoire a=irevoire Fix parts of https://github.com/meilisearch/meilisearch/issues/3318 Co-authored-by: Tamo <tamo@meilisearch.com>	2023-01-11 10:27:13 +00:00
bors[bot]	59704c000c	Merge #3326 3326: Test error codes on settings r=irevoire a=irevoire Co-authored-by: Tamo <tamo@meilisearch.com>	2023-01-11 10:07:52 +00:00
bors[bot]	b117c688f5	Merge #3328 3328: Replace published by released r=Kerollmops a=curquiza Fix a bug introduced here: https://github.com/meilisearch/meilisearch/pull/3229 Regarding this line: > * In multiple CIs: replace the `released` type by `published`, see [here](https://stackoverflow.com/questions/59319281/github-action-different-between-release-created-and-published) why. Will not impact anything, but will prevent to fail our future automation I made mistakes by replacing some un-relevant lines in the - latest git workflow - APT and brew workflow -> the consequence was the workflow ran when releasing `rc0` but they shouldn't have. Luckily the check inside the workflow prevent any release. <img width="1366" alt="Capture d’écran 2023-01-11 à 10 36 52" src="https://user-images.githubusercontent.com/20380692/211771382-d716ff16-0d53-41a9-90de-0d93e01e45fa.png"> This fix is not mandatory thanks to the check inside the workflow, but I would rather roll back to avoid any issues when releasing the official v1 release. Co-authored-by: curquiza <clementine@meilisearch.com>	2023-01-11 09:43:42 +00:00
curquiza	5ec85b7dfb	Replace published by released	2023-01-11 10:30:18 +01:00
bors[bot]	d80be0c28d	Merge #3322 3322: Update mini-dashboard to v0.2.5 r=curquiza a=mdubus Co-authored-by: Morgane Dubus <30866152+mdubus@users.noreply.github.com>	2023-01-11 09:08:11 +00:00
Tamo	398c0c32cd	test all the error codes that can be throw in the settings	2023-01-10 18:19:27 +01:00
Tamo	d4157c0ce4	add a test on the search route for each possible error codes snapshot the json directly instead of using the debug formatting	2023-01-10 17:59:24 +01:00
bors[bot]	98dffbf213	Merge #3317 3317: Remove the unused error codes r=irevoire a=irevoire Remove some unused error code + fix the usage of the search+settings sort and filter error_code Co-authored-by: Tamo <tamo@meilisearch.com>	2023-01-10 16:36:11 +00:00
bors[bot]	11ee7daa0f	Merge #760 760: Add Index::map_size r=Kerollmops a=dureuill # Pull Request ## Related issue Related to discussion: https://github.com/meilisearch/meilisearch/discussions/3280 ## What does this PR do? - Expose `heed::Env::map_size` through `Index::map_size`. This allows knowing after the fact with which `map_size` an environment was opened (which is not always the `map_size` that was configured for the opening of the environment, see the documentation for `Index::map_size`), which will be necessary to guarantee we can reopen the index with a larger `map_size`. ## PR checklist Please check if your PR fulfills the following requirements: - [ ] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [ ] Have you read the contributing guidelines? - [ ] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-01-10 15:59:05 +00:00
Morgane Dubus	f63fee5e97	Update Cargo.toml	2023-01-10 15:11:25 +01:00
Tamo	f0d408c295	update the system http error code to return an internal server error	2023-01-10 14:33:46 +01:00
Tamo	d308684395	remove two ununsed error codes + fix the sort error_code	2023-01-10 11:32:11 +01:00
Louis Dureuil	00746b32c0	Add Index::map_size	2023-01-10 11:16:51 +01:00
bors[bot]	e27bb8ab3e	Merge #3246 3246: Implement most of the error handling enhancement planned for v1.0 r=irevoire a=irevoire Fix #3095 and #2325 Close https://github.com/meilisearch/meilisearch/pull/2540 Implements most of https://github.com/meilisearch/specifications/pull/212 ## Generic error message we re-implements (in deserr): - [x] Json - [x] Incorrect value kind - [x] Missing field - [x] Unknown key - [x] Unexpected - [x] Reimplement the way we show the location - [x] Query parameter - [x] Incorrect value kind - [x] Missing field - [x] Unknown key - [x] Unexpected ## Routes to implements: - [x] Get search - [x] Post search - [x] Settings - [x] Swap indexes - [x] Task API - [x] Documents ressource Error codes to implements; ## Swap API - [x] `duplicate_index_found` → `invalid_swap_duplicate_index_found` ## Search API - [x] `invalid_search_q` - [x] `invalid_search_offset` - [x] `invalid_search_limit` - [x] `invalid_search_page` - [x] `invalid_search_hits_per_page` - [x] `invalid_search_attributes_to_retrieve` - [x] `invalid_search_attributes_to_crop` - [x] `invalid_search_crop_length` - [x] `invalid_search_attributes_to_highlight` - [x] `invalid_search_show_matches_position` - [x] `invalid_search_filter` - [x] `invalid_search_sort` - [x] `invalid_search_facets` - [x] `invalid_search_highlight_pre_tag` - [x] `invalid_search_highlight_post_tag` - [x] `invalid_search_crop_marker` - [x] `invalid_search_matching_strategy` ## Settings API - [x] invalid_settings_displayed_attributes - [x] invalid_settings_searchable_attributes - [x] invalid_settings_filterable_attributes - [x] invalid_settings_sortable_attributes - [x] invalid_settings_ranking_rules - [x] invalid_settings_stop_words - [x] invalid_settings_synonyms - [x] invalid_settings_distinct_attribute - [x] Add invalid_settings_typo_tolerance - [x] ~~invalid_settings_typo_tolerance_min_word_size_for_typos~~ (Merge in invalid_settings_typo_tolerance) - [x] invalid_settings_faceting - [x] invalid_settings_pagination ## Task API - [x] invalid_task_date_filer → invalid_task_before_enqueued_at_filter (for all date filter) ? ## Document Resource - [x] ~~`primary_key_inference_failed` → `index_primary_key_`~~ This doesn't exists anymore after `@dureuill` PR's on the primary key inference ------------------ # Changes # `code` property ## Swap API - [x] `invalid_swap_duplicate_index_found` ✅ [RENAME] - [x] `invalid_swap_indexes` ✅ [NEW] ## Index API ### POST - [x] `missing_index_uid` ✅ [NEW] ### POST/PATCH - [x] `invalid_index_primary_key` ✅ [NEW] ### GET - [x] `invalid_index_limit` ✅ [NEW] - [x] `invalid_index_offset` ✅ [NEW] ## Documents API ### GET - [x] `fields` parameter error `bad_request` → `invalid_document_fields` ✅ [NEW] - [x] `limit` parameter error `bad_request` → `invalid_document_limit` ✅ [NEW] - [x] `offset` parameter error `bad_request` → `invalid_document_offset` ✅ [NEW] ### POST/PUT - [x] `?primaryKey` parameter error `bad_request` → `invalid_index_primary_key` ✅ [NEW] ## Keys API ### POST - ~~`missing_parameter`~~ - [x] `missing_api_key_actions` ✅ [NEW] - [x] `missing_api_key_indexes` ✅ [NEW] - [x] `missing_api_key_expires_at` ✅ [NEW] ### GET - [x] `limit` parameter `bad_request` → `invalid_api_key_limit` ✅ [NEW] - [x] `offset` parameter `bad_request` → `invalid_api_key_offset` ✅ [NEW] ## Misc - [x] ~~`invalid_geo_field`~~ → `invalid_document_geo_field` ✅ [RENAME] # `type` property ## `system` ✅ [NEW] - [x] `no_space_left_on_device` error code - [x] `io_error` error code (does not exist in the current spec, need a catch-up) - [x] `too_many_open_files` error code (does not exist in the current spec, need a catch-up) Co-authored-by: Tamo <tamo@meilisearch.com> Co-authored-by: Loïc Lecrenier <loic.lecrenier@me.com>	2023-01-09 16:25:48 +00:00
Tamo	ff843881c5	remove the documentation of the query parameter extractor module	2023-01-09 15:14:48 +01:00
Loïc Lecrenier	ae08fba76e	Remove forgotten comment	2023-01-09 13:45:03 +01:00
Loïc Lecrenier	af6d4b3031	Remove unused deserr extractor	2023-01-09 13:43:16 +01:00
Louis Dureuil	1cce613399	Fixup dumps-destination -> dump-directory section header in help link	2023-01-09 13:31:57 +01:00
Tamo	b03ee54fe0	makes clippy turbo-happy	2023-01-09 13:04:31 +01:00
Tamo	d17efb9ed6	use the published version of deserr	2023-01-09 12:51:10 +01:00
Loïc Lecrenier	9ab791bedc	Update error codes on the api key routes	2023-01-09 12:30:25 +01:00
Loïc Lecrenier	96105a5e8d	Update error codes on the documents/ routes	2023-01-09 12:30:25 +01:00
Tamo	e706628bb1	fix the error code of the swap index route	2023-01-06 14:48:25 +01:00
Tamo	3c630891bb	fix the error code for the swap index	2023-01-05 21:25:20 +01:00
Tamo	97854274b4	rename the invalid_geo_field error code to invalid_document_geo_field	2023-01-05 21:08:19 +01:00
Tamo	0646f63404	implement the new type property for the system error	2023-01-05 21:06:50 +01:00
Tamo	ce3e8794a2	fix the tests after the rebase	2023-01-05 20:52:26 +01:00
Tamo	50ce0409bc	Integrate deserr on the most important routes	2023-01-05 20:48:29 +01:00
bors[bot]	839b05c43d	Merge #3305 3305: Remove hidden but usable CLI arguments r=Kerollmops a=Kerollmops `@curquiza` found out that we were exposing some internal CLI arguments: `nb-max-chunks` and `log-every-n`. In this PR I removed those two, the only two ones that I found. Those options shouldn't be accessible as non-documented in the documentation or the `--help` message. Fixes https://github.com/meilisearch/meilisearch/issues/3307 Co-authored-by: Clément Renault <clement@meilisearch.com>	2023-01-05 17:11:58 +00:00
bors[bot]	cc699fae40	Merge #3308 3308: Remove `--generate-master-key` option r=Kerollmops a=dureuill # Pull Request ## Related issue Related to https://github.com/meilisearch/specifications/pull/210#issuecomment-1372035525 ## What does this PR do? - Remove the short-lived `--generate-master-key` flag that was too beautiful for this world :D. Removal of this option proceeds of the following reasoning: 1. It is the only option that starts meilisearch and then immediately exits 2. We are unsure if we want to keep it under this form in the future or switch to a subcommand. 3. Releasing this option in v1 would make it insta-stable. 5. The option is only marginally useful, as users will be presented with freshly generated key directly in the error messages if their master key is absent/too short. 6. If we remove this option now, we can still add it back in a future v1 release. If we add it now, we won't be able to remove it in any future v1 version. ## PR checklist Please check if your PR fulfills the following requirements: - [ ] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [ ] Have you read the contributing guidelines? - [ ] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! ### Impacts this impacts the docs team as they would previously have had to document this option, and they may have wanted to use it in the user workflow. Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-01-05 16:19:40 +00:00
Clément Renault	aa4b813237	Derive Default on IndexerOpts	2023-01-05 16:00:45 +01:00
Louis Dureuil	eb08a0fb0b	Remove --generate-master-key option	2023-01-05 14:55:24 +01:00
Clément Renault	cda529c07b	Remove hidden but usable CLI arguments	2023-01-05 14:25:41 +01:00
bors[bot]	1f8ddb366c	Merge #3302 3302: Update insta snap tests for index dates of dump v5 r=curquiza a=loiclec This PR simply updates the content of the insta snapshot test following https://github.com/meilisearch/meilisearch/pull/3013 . I manually verified that the dates in the snaps are indeed correct. Co-authored-by: Loïc Lecrenier <loic.lecrenier@me.com>	2023-01-05 12:58:10 +00:00
bors[bot]	8a3da0c2a7	Merge #3304 3304: Fix update cargo.toml workflow r=Kerollmops a=curquiza Following https://github.com/meilisearch/meilisearch/pull/3224 Fixes #3219 Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com>	2023-01-05 12:16:57 +00:00
Clémentine Urquizar - curqui	c840d55e89	Fix update cargo.toml workflow	2023-01-05 12:56:02 +01:00
bors[bot]	c7a3992510	Merge #3303 3303: Update version for the next release (v1.0.0) in Cargo.toml files r=curquiza a=meili-bot ⚠️ This PR is automatically generated. Check the new version is the expected one before merging. Co-authored-by: curquiza <curquiza@users.noreply.github.com>	2023-01-05 11:53:09 +00:00
curquiza	28408816ef	Update version for the next release (v1.0.0) in Cargo.toml files	2023-01-05 11:45:15 +00:00
bors[bot]	0eaa8ca255	Merge #3266 3266: Improve the way we receive the documents payload- serde multiple ndjson fix r=curquiza a=jiangbo212 # Pull Request ## Related issue Fixes #3037 ## Related PR #3164 ## What does this PR do? Sorry, This PR is mainly to fix the problems caused by my previously provided PR #3164. It causes multiple ndjson data deserialization failures - Fix serde multiple ndjson data failures and add test to it - Fix serde jsonarray error and againest serde it use `from_slice`. only use `from_slice` when serde error category is `data`, it indicate json data is a single json. ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: jiangbo212 <peiyaoliukuan@126.com>	2023-01-05 11:30:29 +00:00
bors[bot]	201bc633d2	Merge #3288 3288: Replace underscores with hyphens in documentation link to error code r=dureuill a=loiclec # Pull Request ## Related issue Fixes #3097 ## Implementation Add a new dependency to `convert_case` (already used transitively by `deserr`) so that the link can be generated using: ```rust /// return the doc url associated with the error fn url(&self) -> String { format!( "https://docs.meilisearch.com/errors#{}", self.name().to_case(convert_case::Case::Kebab) ) } ``` ## Review I'd like the reviewer to check whether it is expected that the content of some `dump` snapshot tests changed :-) Co-authored-by: ManyTheFish <many@meilisearch.com> Co-authored-by: bors[bot] <26634292+bors[bot]@users.noreply.github.com> Co-authored-by: Loïc Lecrenier <loic.lecrenier@me.com>	2023-01-05 11:08:57 +00:00
Loïc Lecrenier	ba839852f5	Update insta snap tests for index dates of dump v5	2023-01-05 11:45:40 +01:00
Louis Dureuil	be9786bed9	Change primary key inference error messages	2023-01-05 10:40:09 +01:00
Loïc Lecrenier	f9aa897ab5	Update insta tests	2023-01-05 10:19:19 +01:00
Loïc Lecrenier	2d74678b51	Replace underscores with hyphens in doc link to error code	2023-01-05 10:09:02 +01:00
bors[bot]	db7eaf23f4	Merge #3251 3251: Add a specific test on finite pagination placeolder search with disti… r=curquiza a=ManyTheFish Add a specific test on finite pagination placeholder search with distinct attributes related to https://github.com/meilisearch/milli/pull/743 related to https://github.com/meilisearch/meilisearch/issues/3200 poke `@curquiza` > note that the destination branch should be changed Co-authored-by: ManyTheFish <many@meilisearch.com>	2023-01-05 09:06:53 +00:00
bors[bot]	32f7cfa5cb	Merge #3295 3295: Adjust Master Key-related messages r=dureuill a=dureuill # Pull Request ## Related issue Follow up for #3272 ## What does this PR do? - Consistently capitalize "master key" (instead of "Master Key" sometimes) (see https://github.com/meilisearch/specifications/pull/209#discussion_r1060081094) - Clarify that the counted unit for master key length is bytes, not characters (see https://github.com/meilisearch/documentation/issues/2069#issuecomment-1368873167) ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-01-05 08:43:23 +00:00
bors[bot]	a402fc4486	Merge #3013 3013: Extract the dates out of the dumpv5. r=loiclec a=funilrys Hi there, please review this PR that tries to fix #2986. I'm still learning Rust and I found that #2986 is an excellent way for me to read and learn what others do with Rust. So please excuse my semantics ... Stay safe and healthy. --- # Pull Request This patch possibly fixes #2986. This patch introduces a way to fill the IndexMetadata.created_at and IndexMetadata.updated_at keys from the tasks events. This is done by reading the creation date of the first event (created_at) and the creation date of the last event (updated_at). ## Related issue Fixes #2986 ## What does this PR do? - Extract the dates out of the dumpv5. ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: funilrys <contact@funilrys.com>	2023-01-05 08:23:52 +00:00
bors[bot]	502d9e4b24	Merge #3278 3278: Remove `--max-index-size` and `--max-task-db-size` flags r=Kerollmops a=dureuill # Pull Request ## Related issue Fixes #3231 ## What does this PR do? - Remove `--max-index-size` and `--max-task-db-size` flags from the CLI, config file and environment variable - Set the size of all indexes to 500GiB and the size of the task DB to 10GiB. Reviewers might want to review these values carefully. ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-01-04 16:44:27 +00:00
Louis Dureuil	a85ff1f690	Fix documentation Co-authored-by: Clément Renault <clement@meilisearch.com>	2023-01-04 17:20:03 +01:00
Louis Dureuil	233372abea	Remove `--max-index-size` and `--max-task-db-size`	2023-01-04 17:20:01 +01:00
bors[bot]	13d4ae264a	Merge #3269 3269: Simplify primary key inference r=dureuill a=dureuill # Pull Request ## Related issue Related to https://github.com/meilisearch/meilisearch/issues/3233 ## What does this PR do? - Integrates https://github.com/meilisearch/milli/pull/752 in meilisearch - Remove `Serialize` and `Deserialize` from `error::Code` as it is unused. - No longer filter on `milli` logs when `--log-level` is "info". - `milli` only has the newly-added inference log at the `info` level (from greping `info` in the codebase) - the default value for `--log-level` is "INFO" and not "info" since `v0.30` so the filter is not active by default. - updates milli to v0.38.0 ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-01-04 16:14:36 +00:00
bors[bot]	c766e06003	Merge #3281 3281: Merge `--schedule-snapshot` and `--snapshot-interval-sec` options r=dureuill a=dureuill # Pull Request ## Related issue Fixes #3131 ## What does this PR do? - Removes `--snapshot-interval-sec` - `--schedule-snapshot` now accepts an optional integer value specifying the interval in seconds - The config file no longer has a snapshot_interval_sec key. Instead, the schedule_snapshot key now additionally accepts an integer value specifying the interval in seconds - The env variable MEILI_SNAPSHOT_INTERVAL no longer exists - The env variable MEILI_SCHEDULE_SNAPSHOT is always specified to the interval of the snapshot in seconds when defined. If snapshots are disabled the variable is undefined. --- Relevant part of the `--help` <img width="885" alt="Capture d’écran 2022-12-27 à 18 22 32" src="https://user-images.githubusercontent.com/41078892/209700626-1a1292c1-14e3-45b6-8265-e0adbd76ecf1.png"> --- ### Tests \| `schedule_snapshot` in config.toml \| `--schedule-snapshot` flag on CLI \| `MEILI_SCHEDULE_SNAPSHOT` \| `opt.schedule_snapshot` \| \|--\|--\|--\|--\| \| missing \| missing \| missing \| `Disabled` \| `false` \| missing \| missing \| `Disabled` \| `true` \| missing \| missing \| `Enabled(86400)` \| `1234` \| missing \| missing \| `Enabled(1234)` \| missing \| `--schedule-snapshot` \| missing \| `Enabled(86400)` \| `false` \| `--schedule-snapshot` \| missing \| `Enabled(86400)` \| missing \| `--schedule-snapshot 2345` \| missing \| `Enabled(2345)` \| `false` \| `--schedule-snapshot 2345` \| missing \| `Enabled(2345)` \| `true` \| `--schedule-snapshot 2345` \| missing \| `Enabled(2345)` \| `1234` \| `--schedule-snapshot 2345` \| missing \| `Enabled(2345)` \| `false` \| `--schedule-snapshot 2345` \| 3456 \| `Enabled(2345)` \| `false` \| `--schedule-snapshot` \| 3456 \| `Enabled(86400)` \| `1234` \| missing \| 3456 \| `Enabled(3456)` \| `false` \| missing \| 3456 \| `Enabled(3456)` ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-01-04 14:25:47 +00:00
Louis Dureuil	fcbd47281b	Fix tests	2023-01-04 14:24:20 +01:00
Louis Dureuil	b6d80293f7	Propagate new error codes from milli	2023-01-04 14:24:20 +01:00
Louis Dureuil	0e98a71a24	Update milli to v0.38	2023-01-04 14:24:20 +01:00
Louis Dureuil	5cb566b165	No longer filter out milli logs when --log-level is "info"	2023-01-04 14:24:20 +01:00
Louis Dureuil	9d46caba29	Code doesn't need to be serializable/deserializable	2023-01-04 14:16:22 +01:00
Louis Dureuil	c4aa5cc7d0	Merge --schedule-snapshot and --snapshot-interval-sec options	2023-01-04 14:13:54 +01:00
bors[bot]	12c3d432f9	Merge #3293 3293: Explicitly restrict log level options to those that are documented r=loiclec a=loiclec Fixes https://github.com/meilisearch/meilisearch/issues/3292 Co-authored-by: Loïc Lecrenier <loic.lecrenier@me.com>	2023-01-04 10:30:35 +00:00
bors[bot]	c3f4835e8e	Merge #733 733: Avoid a prefix-related worst-case scenario in the proximity criterion r=loiclec a=loiclec # Pull Request ## Related issue Somewhat fixes (until merged into meilisearch) https://github.com/meilisearch/meilisearch/issues/3118 ## What does this PR do? When a query ends with a word and a prefix, such as: ``` word pr ``` Then we first determine whether `pre` could possibly be in the proximity prefix database before querying it. There are then three possibilities: 1. `pr` is not in any prefix cache because it is not the prefix of many words. We don't query the proximity prefix database. Instead, we list all the word derivations of `pre` through the FST and query the regular proximity databases. 2. `pr` is in the prefix cache but cannot be found in the proximity prefix databases. In this case, we partially disable the proximity ranking rule for the pair `word pre`. This is done as follows: 1. Only find the documents where `word` is in proximity to `pre` exactly (no derivations) 2. Otherwise, assume that their proximity in all the documents in which they coexist is >= 8 3. `pr` is in the prefix cache and can be found in the proximity prefix databases. In this case we simply query the proximity prefix databases. Note that if a prefix is longer than 2 bytes, then it cannot be in the proximity prefix databases. Also, proximities larger than 4 are not present in these databases either. Therefore, the impact on relevancy is: 1. For common prefixes of one or two letters: we no longer distinguish between proximities from 4 to 8 2. For common prefixes of more than two letters: we no longer distinguish between any proximities 3. For uncommon prefixes: nothing changes Regarding (1), it means that these two documents would be considered equally relevant according to the proximity rule for the query `heard pr` (IF `pr` is the prefix of more than 200 words in the dataset): ```json [ { "text": "I heard there is a faster proximity criterion" }, { "text": "I heard there is a faster but less relevant proximity criterion" } ] ``` Regarding (2), it means that two documents would be considered equally relevant according to the proximity rule for the query "faster pro": ```json [ { "text": "I heard there is a faster but less relevant proximity criterion" } { "text": "I heard there is a faster proximity criterion" }, ] ``` But the following document would be considered more relevant than the two documents above: ```json { "text": "I heard there is a faster swimmer who is competing in the pro section of the competition " } ``` Note, however, that this change of behaviour only occurs when using the set-based version of the proximity criterion. In cases where there are fewer than 1000 candidate documents when the proximity criterion is called, this PR does not change anything. --- ## Performance I couldn't use the existing search benchmarks to measure the impact of the PR, but I did some manual tests with the `songs` benchmark dataset. ``` 1. 10x 'a': - 640ms ⟹ 630ms = no significant difference 2. 10x 'b': - set-based: 4.47s ⟹ 7.42 = bad, ~2x regression - dynamic: 1s ⟹ 870 ms = no significant difference 3. 'Someone I l': - set-based: 250ms ⟹ 12 ms = very good, x20 speedup - dynamic: 21ms ⟹ 11 ms = good, x2 speedup 4. 'billie e': - set-based: 623ms ⟹ 2ms = very good, x300 speedup - dynamic: ~4ms ⟹ 4ms = no difference 5. 'billie ei': - set-based: 57ms ⟹ 20ms = good, ~2x speedup - dynamic: ~4ms ⟹ ~2ms. = no significant difference 6. 'i am getting o' - set-based: 300ms ⟹ 60ms = very good, 5x speedup - dynamic: 30ms ⟹ 6ms = very good, 5x speedup 7. 'prologue 1 a 1: - set-based: 3.36s ⟹ 120ms = very good, 30x speedup - dynamic: 200ms ⟹ 30ms = very good, 6x speedup 8. 'prologue 1 a 10': - set-based: 590ms ⟹ 18ms = very good, 30x speedup - dynamic: 82ms ⟹ 35ms = good, ~2x speedup ``` Performance is often significantly better, but there is also one regression in the set-based implementation with the query `b b b b b b b b b b`. Co-authored-by: Loïc Lecrenier <loic.lecrenier@me.com>	2023-01-04 09:00:50 +00:00
Loïc Lecrenier	d082ded7ad	Explicitly restrict log level options to those that are documented Fixes https://github.com/meilisearch/meilisearch/issues/3292	2023-01-04 09:40:24 +01:00
bors[bot]	49f58b2c47	Merge #732 732: Interpret synonyms as phrases r=loiclec a=loiclec # Pull Request ## Related issue Fixes (when merged into meilisearch) https://github.com/meilisearch/meilisearch/issues/3125 ## What does this PR do? We now map multi-word synonyms to phrases instead of loose words. Such that the request: ``` btw I am going to nyc soon ``` is interpreted as (when the synonym interpretation is chosen for both `btw` and `nyc`): ``` "by the way" I am going to "New York City" soon ``` instead of: ``` by the way I am going to New York City soon ``` This prevents queries containing multi-word synonyms to exceed to word length limit and degrade the search performance. In terms of relevancy, there is a debate to have. I personally think this could be considered an improvement, since it would be strange for a user to search for: ``` good DIY project ``` and have a result such as: ``` { "text": "whether it is a good project to do, you'll have to decide for yourself" } ``` However, for synonyms such as `NYC -> New York City`, then we will stop matching documents where `New York` is separated from `City`. This is however solvable by adding an additional mapping: `NYC -> New York`. ## Performance With the old behaviour, some long search requests making heavy uses of synonyms could take minutes to be executed. This is no longer the case, these search requests now take an average amount of time to be resolved. Co-authored-by: Loïc Lecrenier <loic.lecrenier@me.com>	2023-01-04 08:34:18 +00:00
bors[bot]	947f08793a	Merge #3296 3296: Remove `--disable-auto-batching` CLI option r=gmourier a=loiclec Fixes #3294 The `index-scheduler` code is not modified, only the CLI options have changed. Co-authored-by: Loïc Lecrenier <loic.lecrenier@me.com>	2023-01-03 16:57:14 +00:00
bors[bot]	6a10e85707	Merge #736 736: Update charabia r=curquiza a=ManyTheFish Update Charabia to the last version. > We are now Romanizing Chinese characters into Pinyin. > Note that we keep the accent because they are in fact never typed directly by the end-user, moreover, changing an accent leads to a different Chinese character, and I don't have sufficient knowledge to forecast the impact of removing accents in this context. Co-authored-by: ManyTheFish <many@meilisearch.com>	2023-01-03 15:44:41 +00:00
Louis Dureuil	17dac72464	Characters -> bytes	2023-01-03 15:31:02 +01:00
Loïc Lecrenier	b821c72459	Remove `--disable-auto-batching` CLI option	2023-01-03 15:01:04 +01:00
Louis Dureuil	7b2575c646	Master Key -> master key	2023-01-03 14:45:23 +01:00
bors[bot]	c505fa9d7d	Merge #758 758: Bump taiki-e/install-action from 1 to 2 r=curquiza a=dependabot[bot] Bumps [taiki-e/install-action](https://github.com/taiki-e/install-action) from 1 to 2. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/taiki-e/install-action/releases">taiki-e/install-action's releases</a>.</em></p> <blockquote> <h2>2.0.0</h2> <p>This release implements a mechanism to automatically track the latest version of the tool on our end. (<a href="https://github-redirect.dependabot.com/taiki-e/install-action/pull/27">#27</a>) Hopefully, this will avoid situations such as "new version of the tool has been released, but the maintainer has not been aware of it for a number of months". This also makes it easier to add support for new tools.</p> <p>This release also includes the following improvements:</p> <ul> <li> <p>Verify SHA256 checksums for downloaded files in all tools installed from GH releases. (<a href="https://github-redirect.dependabot.com/taiki-e/install-action/pull/27">#27</a>)</p> </li> <li> <p>Support omitting the patch/minor version in all tools installed from GH releases. (<a href="https://github-redirect.dependabot.com/taiki-e/install-action/pull/27">#27</a>)</p> <p>For example:</p> <pre lang="yaml"><code>- uses: taiki-e/install-action@v2 with: tool: cargo-hack@0.5 </code></pre> <p>You can also omit the minor version if the major version of tool is 1 or greater.</p> </li> <li> <p>Support <code>just</code>. (<a href="https://github-redirect.dependabot.com/taiki-e/install-action/pull/34">#34</a>)</p> </li> <li> <p>Support <code>dprint</code>. (<a href="https://github-redirect.dependabot.com/taiki-e/install-action/pull/34">#34</a>)</p> </li> </ul> <p>Note: This release is considered a breaking change because installing on versions not yet recognized by the action or on pre-release versions will no longer work with this release. (They were never officially supported, but they could work before.) Please submit an issue if you need these supports again.</p> <h2>1.17.3</h2> <ul> <li>Update <code>wasmtime@latest</code> to 4.0.0.</li> </ul> <h2>1.17.2</h2> <ul> <li>Update <code>mdbook@latest</code> to 0.4.25.</li> </ul> <h2>1.17.1</h2> <ul> <li>Update <code>mdbook@latest</code> to 0.4.23.</li> <li>Support <code>mdbook</code> on Linux (musl).</li> <li>Update <code>cargo-llvm-cov@latest</code> to 0.5.3.</li> </ul> <h2>1.17.0</h2> <ul> <li>Update <code>protoc@latest</code> to 3.21.12.</li> <li>Support aarch64 self-hosted runners (Linux, macOS, Windows).</li> <li>Improve support for Fedora/RHEL based containers/self-hosted runners.</li> </ul> <h2>1.16.0</h2> <ul> <li> <p>Update <code>cargo-binstall@latest</code> to 0.18.1. (<a href="https://github-redirect.dependabot.com/taiki-e/install-action/pull/32">#32</a>, thanks <a href="https://github.com/NobodyXu"><code>`@NobodyXu</code></a>)</p>` </li> <li> <p>If the host environment lacks packages required for installation, such as <code>curl</code> or <code>tar</code>, install them if possible.</p> <p>It is mainly intended to make the use of this action easy on containers or self-hosted runners, and currently supports Debian-based distributions (including Ubuntu) and Alpine.</p> </li> </ul> <!-- raw HTML omitted --> </blockquote> <p>... (truncated)</p> </details> <details> <summary>Changelog</summary> <p><em>Sourced from <a href="https://github.com/taiki-e/install-action/blob/main/CHANGELOG.md">taiki-e/install-action's changelog</a>.</em></p> </details> <details> <summary>Commits</summary> <ul> <li><a href="`8ffc26aecd`"><code>8ffc26a</code></a> Release 2.0.1</li> <li><a href="`82e9eb5996`"><code>82e9eb5</code></a> Update DEVELOPMENT.md</li> <li><a href="`d3b7ad8380`"><code>d3b7ad8</code></a> Update changelog</li> <li><a href="`f1a96ee3ed`"><code>f1a96ee</code></a> Update changelog when manifest update</li> <li><a href="`46063c186c`"><code>46063c1</code></a> Update cargo-minimal-versions</li> <li><a href="`048586d7a8`"><code>048586d</code></a> Update cargo-hack</li> <li><a href="`d117b8d41a`"><code>d117b8d</code></a> Remove outdated todo</li> <li><a href="`76828c33cd`"><code>76828c3</code></a> Release 2.0.0</li> <li><a href="`eea8c318de`"><code>eea8c31</code></a> Update readme and changelog</li> <li><a href="`ab0e193cf5`"><code>ab0e193</code></a> Support dprint</li> <li>Additional commits viewable in <a href="https://github.com/taiki-e/install-action/compare/v1...v2">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=taiki-e/install-action&package-manager=github_actions&previous-version=1&new-version=2)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) You can trigger a rebase of this PR by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) </details> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2023-01-03 10:44:36 +00:00
bors[bot]	ab655a85e8	Merge #3279 3279: Clarify error message when the db and engine versions are incompatible r=irevoire a=dureuill # Pull Request ## Related issue Related to https://github.com/meilisearch/meilisearch/issues/2752 ## What does this PR do? - Implements https://github.com/meilisearch/product/discussions/572#discussioncomment-4390616 ## PR checklist Please check if your PR fulfills the following requirements: - [ ] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-01-02 17:18:11 +00:00
bors[bot]	6425e06cf2	Merge #3274 3274: Reject master keys that are less than 16 bytes and add `--generate-master-key` CLI option r=irevoire a=dureuill # Pull Request ## Related issue Fix #3272 Fix #3287 ## What does this PR do? ### User standpoint --- - Adds a `--generate-master-key` CLI flag to generate a fresh Master Key and exit. <img width="1351" alt="Capture d’écran 2022-12-22 à 14 18 58" src="https://user-images.githubusercontent.com/41078892/209142778-eab52eeb-eaa8-409b-897a-c0d5728c8aaa.png"> --- (relevant fragment of the `--help` message) <img width="1351" alt="Capture d’écran 2022-12-22 à 14 19 40" src="https://user-images.githubusercontent.com/41078892/209142891-ebfa2ed6-f231-4f76-a3ae-b7542c7aef04.png"> --- - When `meilisearch` is started in the `development` environment and no Master Key has been provided, then the binary prints a warning before starting. <img width="1351" alt="Capture d’écran 2022-12-22 à 14 14 49" src="https://user-images.githubusercontent.com/41078892/209142158-54eba3b7-bf71-4f3f-8840-0600b13a1a9f.png"> --- - When `meilisearch` is started in the `development` environment and the provided Master Key is shorter than 16 bytes, then the binary prints a warning before starting. <img width="1351" alt="Capture d’écran 2022-12-22 à 14 15 58" src="https://user-images.githubusercontent.com/41078892/209142295-0209fe47-c03b-424f-a73f-cee9b633137a.png"> --- - When `meilisearch` is started in the `production` environment, and no Master Key is provided, the error message is altered to generate a fresh Master Key. <img width="1351" alt="Capture d’écran 2022-12-22 à 17 29 02" src="https://user-images.githubusercontent.com/41078892/209180540-0def5798-15db-47f0-a6ec-8cfa081dea77.png"> --- - When `meilisearch` is started in the `production` environment, and the provided Master Key is shorter than 16 bytes, then the binary exits with an error. <img width="1351" alt="Capture d’écran 2022-12-22 à 17 28 47" src="https://user-images.githubusercontent.com/41078892/209180567-fa54fe33-fbc4-4b9f-b281-7dfb7b33af85.png"> --- This implements the solution B described here: https://github.com/meilisearch/product/discussions/538#discussioncomment-4391346 ### Implementation standpoint - Add a new `meilisearch-auth::generate_master_key` function that uses a Cryptographic Random Number Generator (CRNG) to fill a vector of 32 bytes before encoding these bytes as base64 ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Louis Dureuil <louis@meilisearch.com> Co-authored-by: Tamo <tamo@meilisearch.com>	2023-01-02 16:00:40 +00:00
Tamo	1692f58b83	slightly update the message associated with the cli parameter + accept an env variable	2023-01-02 16:49:35 +01:00
Tamo	9ba4d0f921	update the error messages according to the spec	2023-01-02 16:43:23 +01:00
Tamo	4b6ffe0cd1	Update meilisearch-auth/src/lib.rs	2023-01-02 16:33:02 +01:00
bors[bot]	336c77aa45	Merge #3245 3245: Enable create_raw_index(...) to specify time r=irevoire a=amab8901 # Pull Request ## Related issue Partially fixes #2983 ## What does this PR do? - Enables [`create_raw_index`](`660be071b5/index-scheduler/src/lib.rs (L868)`) to specify time ## PR checklist Please check if your PR fulfills the following requirements: - [ X ] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [ X ] Have you read the contributing guidelines? - [ X ] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: amab8901 <amab8901@protonmail.com>	2023-01-02 14:11:39 +00:00
bors[bot]	9519e60f97	Merge #709 709: Optimise the `ExactWords` sub-criterion within `Exactness` r=loiclec a=loiclec # Pull Request ## Related issue Fixes (partially) https://github.com/meilisearch/meilisearch/issues/3116 ## What does this PR do? 1. Reduces the algorithmic complexity of finding the documents containing N exact words from something that is exponential to something that is polynomial. 2. Cache intermediary results between different calls to the `exactness` criterion. ## Performance Results On the `smol_songs.csv` dataset, a request containing 10 common words now takes about 60ms instead of 5 seconds to execute. For example, this is the case with this (admittedly nonsensical) request: `Rock You Hip Hop Folk World Country Electronic Love The`. Co-authored-by: Loïc Lecrenier <loic.lecrenier@me.com>	2023-01-02 12:28:30 +00:00
Loïc Lecrenier	b5df889dcb	Apply review suggestions: simplify implementation of exactness criterion	2023-01-02 13:11:47 +01:00
bors[bot]	31155dce4c	Merge #752 752: Simplify primary key inference r=irevoire a=dureuill # Pull Request ## Related issue Related to https://github.com/meilisearch/meilisearch/issues/3233 ## What does this PR do? ### User PoV - Change primary key inference to only consider a value as a candidate when it ends with "id", rather than when it simply contains "id". - Change primary key inference to always fail when there are multiple candidates. - Replace UserError::MissingPrimaryKey with `UserError::NoPrimaryKeyCandidateFound` and `UserError::MultiplePrimaryKeyCandidatesFound` ### Implementation-wise - Remove uses of UserError::MissingPrimaryKey not pertaining to inference. This introduces a possible panicking path. ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2023-01-02 11:44:22 +00:00
Loïc Lecrenier	8d36570958	Add explicit criterion impl strategy to proximity search tests	2023-01-02 10:37:01 +01:00
dependabot[bot]	939e7faf31	Bump taiki-e/install-action from 1 to 2 Bumps [taiki-e/install-action](https://github.com/taiki-e/install-action) from 1 to 2. - [Release notes](https://github.com/taiki-e/install-action/releases) - [Changelog](https://github.com/taiki-e/install-action/blob/main/CHANGELOG.md) - [Commits](https://github.com/taiki-e/install-action/compare/v1...v2) --- updated-dependencies: - dependency-name: taiki-e/install-action dependency-type: direct:production update-type: version-update:semver-major ... Signed-off-by: dependabot[bot] <support@github.com>	2023-01-01 10:02:00 +00:00
bors[bot]	776acb5ed3	Merge #3276 3276: README: Replace Slack link with Discord r=dureuill a=shivaylamba # Pull Request ## Related issue Fixes #3275 ## What does this PR do? Update Slack link with Discord link in the README ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Shivay Lamba <shivaylamba@gmail.com>	2022-12-29 10:32:42 +00:00
Louis Dureuil	3e9834abff	Change error message when the db version is incompatible with engine version.	2022-12-26 17:34:36 +01:00
Louis Dureuil	3cba476a9f	Add `--generate-master-key` CLI option	2022-12-26 10:36:45 +01:00
Louis Dureuil	57e851d8a9	Check for key length	2022-12-26 10:36:45 +01:00
Shivay Lamba	9c45850bd2	README: Replace Slack link with Discord	2022-12-26 00:19:13 +05:30
funilrys	3e0e8164a3	fixup! Adjust + Cleanup changes.	2022-12-22 18:01:54 +01:00
funilrys	0bc4572905	Adjust + Cleanup changes. Indeed, I missed some of the changed that were introduced by #3190.	2022-12-22 17:53:33 +01:00
funilrys	4e6c663a2e	Release unecessary ownership.	2022-12-22 17:47:58 +01:00
funilrys	e2775c6f49	Remove unused object.	2022-12-22 17:47:58 +01:00
funilrys	c07a5932cb	Apply fmt.	2022-12-22 17:47:58 +01:00
funilrys	528a944997	Reimplement v5 date extraction. Indeed, before this patch the implementation wasn't correct.	2022-12-22 17:47:58 +01:00
funilrys	13fb5ce974	Re-Open tasks list when needed. Indeed, before this patch we were using the reference instead of "reopening" the task list each time we needed to access it. Without this patch, all other usage of the task attribute will break.	2022-12-22 17:47:57 +01:00
funilrys	a43a0712fa	Add reader.v5.tasks.Task.updated_at. There was no way to "quickly" get the update date.	2022-12-22 17:47:57 +01:00
funilrys	1be4619b91	Add reader.v5.tasks.Task.created_at. There was no way to "quickly" get the creation date.	2022-12-22 17:47:57 +01:00
funilrys	cf50f85986	Add reader.v5.tasks.Task.processed_at. There was no way to "quickly" get the processed date.	2022-12-22 17:47:57 +01:00
funilrys	61b3a29ff3	Extract the dates out of the dumpv5. This patch possibly fixes #2986. This patch introduces a way to fill the IndexMetadata.created_at and IndexMetadata.updated_at keys from the tasks events. This is done by reading the creation date of the first event (created_at) and the creation date of the last event (updated_at).	2022-12-22 17:47:57 +01:00
Loïc Lecrenier	32c6062e65	Optimise exactness criterion 1. Cache some results between calls to next() 2. Compute the combinations of exact words more efficiently	2022-12-22 12:28:45 +01:00
Loïc Lecrenier	f097aafa1c	Add unit test for prefix handling by the proximity criterion	2022-12-22 12:08:00 +01:00
Loïc Lecrenier	777b387dc4	Avoid a prefix-related worst-case scenario in the proximity criterion	2022-12-22 12:08:00 +01:00
Loïc Lecrenier	b0f3dc2c06	Interpret synonyms as phrases	2022-12-22 12:07:51 +01:00
Louis Dureuil	66e18eae79	auth: add generate_master_key function	2022-12-22 11:55:27 +01:00
amab8901	9a39c4e40d	Get date from IndexMetaData	2022-12-22 11:46:17 +01:00
amab8901	df176aaf01	Insert dump_reader.date() into create_raw_index(_) argument	2022-12-21 15:16:31 +01:00
Louis Dureuil	4b166bea2b	Add primary_key_inference test	2022-12-21 15:13:38 +01:00
Louis Dureuil	5943100754	Fix existing tests	2022-12-21 15:13:38 +01:00
Louis Dureuil	b24def3281	Add logging when inference took place. Displays log message in the form: ``` [2022-12-21T09:19:42Z INFO milli::update::index_documents::enrich] Primary key was not specified in index. Inferred to 'id' ```	2022-12-21 15:13:38 +01:00
Louis Dureuil	402dcd6b2f	Simplify primary key inference	2022-12-21 15:13:38 +01:00
Louis Dureuil	13c95d25aa	Remove uses of UserError::MissingPrimaryKey not related to inference	2022-12-21 15:13:36 +01:00
amab8901	0893b175dc	Merge branch 'main' into 2983-forward-date-to-milli	2022-12-21 14:31:19 +01:00
amab8901	d5978d11e1	Refactor	2022-12-21 14:28:00 +01:00
bors[bot]	a8defb585b	Merge #742 742: Add a "Criterion implementation strategy" parameter to Search r=irevoire a=loiclec Add a parameter to search requests which determines the implementation strategy of the criteria. This can be either `set-based`, `iterative`, or `dynamic` (ie choosing between set-based or iterative at search time). See https://github.com/meilisearch/milli/issues/755 for more context about this change. Co-authored-by: Loïc Lecrenier <loic.lecrenier@me.com>	2022-12-21 12:18:49 +00:00
Loïc Lecrenier	339a4b0789	Make clippy happy	2022-12-21 12:49:34 +01:00
Loïc Lecrenier	904fd2f6d1	Add a search strategy option to the cli	2022-12-21 12:48:53 +01:00
Loïc Lecrenier	229405aeb9	Choose implementation strategy of criterion at runtime	2022-12-21 09:29:39 +01:00
jiangbo212	2780e365e2	test update and ndjson serde use from_slice	2022-12-21 14:31:45 +08:00
jiangbo212	bf2a401a05	serde ndjson fix	2022-12-21 11:27:15 +08:00
bors[bot]	9925309492	Merge #3263 3263: Handle most io error instead of tagging everything as an internal r=dureuill a=irevoire Fix https://github.com/meilisearch/meilisearch/issues/2255 Fix https://github.com/meilisearch/meilisearch/issues/2785 Close https://github.com/meilisearch/milli/pull/580 - [x] Find a way to catch the `io::Error` contained in `serde_json::Error`: We can't: https://docs.rs/serde_json/latest/serde_json/struct.Error.html - [x] Check the `grenad::Error` as well => the `grenad::Error::Io` error are correctly converted to a `milli::Error::Io` error - [x] Ensure the error code mean the same thing under windows Co-authored-by: Tamo <tamo@meilisearch.com>	2022-12-20 17:15:53 +00:00
Tamo	9e0cce5ca4	Update dump/src/error.rs Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2022-12-20 18:08:51 +01:00
Tamo	336ea57384	Update dump/src/error.rs Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2022-12-20 18:08:44 +01:00
Tamo	c637bfba37	convert all the document format error due to io to io::Error	2022-12-20 17:49:38 +01:00
Tamo	3040172562	update the error message as well	2022-12-20 17:31:13 +01:00
bors[bot]	249e051cd4	Merge #750 750: Fix hard-deletion of an external id that was soft-deleted and then reimported - main r=irevoire a=loiclec # Pull Request ## Related issue Fixes (when merged into meilisearch) https://github.com/meilisearch/meilisearch/issues/3021 ## What does this PR do? There was a bug happening when: 1. Documents were added 2. Some of these documents were replaced using soft-deletion 3. A deletion of another non-replaced document takes place and triggers a hard-deletion 4. Documents with the same identifiers as the replaced documents are added again Then, search results would return duplicate documents. No crash would happen at any time (this is the reason it wasn't caught by the previous fuzz test. I have updated the new one such that it also checks the result of a placeholder search request, which then finds the bug immediately). The cause of the bug is: 1. When a hard-deletion is triggered, we try to retrieve the external document id associated with each soft-deleted document id. 2. Then, we take this list of external document ids and remove each of them from the `ExternalDocumentsIds` structure. 3. However, this is not correct in case an existing (non-deleted) document shares the external id of a soft-deleted document. ## Implementation of the fix 1. Before we process a permanent deletion, we update the list of soft-deleted document ids. 2. Then, the permanent deletion's job is to remove the soft-deleted documents from all data structures. Therefore, to update `ExternalDocumentsIds`, we can simply call the `delete_soft_deleted_documents_ids_from_fsts` method, which is faster and simpler. ## Correctness A unit test was added to reproduce the bug. The new fuzz test, when adjusted to check the correctness of a placeholder search, could also instantly reproduce the bug, but now does not find any other problem. Co-authored-by: Loïc Lecrenier <loic.lecrenier@me.com>	2022-12-20 16:13:20 +00:00
Tamo	52aa34d984	remove an unused error handling file	2022-12-20 16:32:51 +01:00
Loïc Lecrenier	fc0e7382fe	Fix hard-deletion of an external id that was soft-deleted	2022-12-20 15:33:31 +01:00
bors[bot]	2c86d42a44	Merge #3264 3264: Remove macos-latest and windows-latest usages r=curquiza a=curquiza Related to https://github.com/meilisearch/meilisearch/issues/3109#issuecomment-1359151297 Remove the `macos-latest` and `windows-latest` to replace them with the specific version: this will avoid "surprises" in the future when GitHub changes the `latest` version. This way, it will also allow us to let the documentation team know about the changes, since we will control the macOS/Windows version we support Co-authored-by: curquiza <clementine@meilisearch.com>	2022-12-20 10:53:37 +00:00
curquiza	8ce3a34ffa	Remove macos-latest and windows-latest usages	2022-12-20 11:10:09 +01:00
bors[bot]	259c04eb28	Merge #3261 3261: Use ubuntu-18.04 container instead of GitHub hosted actions r=curquiza a=curquiza Related to (but does not fix totally) https://github.com/meilisearch/meilisearch/issues/3109 and https://github.com/meilisearch/product/discussions/547#discussioncomment-4109143 ## For reviewers, what's the PR changes: - Use ubuntu-latest where compiling with ubuntu-18.04 is not needed (`update-version-cargo-toml`, `fmt`, `clippy` jobs) - Where ubuntu-18.04 is required - Use `ubuntu-latest` as runner - Use `ubuntu:18.04` as Docker container - Install the required dependencies (curl and cc) - Use `actions-rs/toolchain@v1` instead of `hecrj/setup-rust-action@master`. It's more stable and followed alternative. Plus it was easy to make it work with our container contrary to the old one. Change applied in all our CIs to be more consistent - Remove some useless space to increase readability. Co-authored-by: curquiza <clementine@meilisearch.com>	2022-12-20 09:28:09 +00:00
Tamo	d8fb506c92	handle most io error instead of tagging everything as an internal	2022-12-19 20:50:40 +01:00
amab8901	aa03e02fdc	Apply Rustfmt	2022-12-19 19:24:56 +01:00
curquiza	7ef23addb6	Add comment to bring more context	2022-12-19 18:46:27 +01:00
bors[bot]	97fb64e40e	Merge #747 747: Soft-deletion computation no longer depends on the mapsize r=irevoire a=dureuill # Pull Request ## Related issue Related to https://github.com/meilisearch/meilisearch/issues/3231: After removing `--max-index-size`, the `mapsize` will always be unrelated to the actual max size the user wants for their DB, so it doesn't make sense to use these values any longer. This implements solution 2.3 from https://github.com/meilisearch/meilisearch/issues/3231#issuecomment-1348628824 ## What does this PR do? ### User-visible - Soft-deleted are no longer deleted when there is less than 10% of the mapsize available or when they take more than 10% of the mapsize - Instead, they are deleted when they are more soft deleted than regular documents, or when they take more than 1GiB disk space (estimated). ### Implementation standpoint 1. Adds a `DeletionStrategy` struct to replace the boolean `disable_soft_deletion` that we had up until now. This enum allows us to specify that we want "always hard", "always soft", or to use the dynamic soft-deletion strategy (default). 2. Uses the current strategy when deleting documents, with the new heuristics being used in the `DeletionStrategy::Dynamic` variant. 3. Updates the tests to use the appropriate DeletionStrategy whenever needed (one of `AlwaysHard` or `AlwaysSoft` depending on the test) Note to reviewers: this PR is optimized for a commit-by-commit review. ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Louis Dureuil <louis@meilisearch.com> Co-authored-by: Tamo <tamo@meilisearch.com>	2022-12-19 17:46:18 +00:00
curquiza	b3fce7c366	Remove useless continue-on-error	2022-12-19 18:39:35 +01:00
curquiza	5099a40484	Use ubuntu-18.04 container in publish CIs	2022-12-19 18:35:33 +01:00
Tamo	69edbf9f6d	Update milli/src/update/delete_documents.rs	2022-12-19 18:23:50 +01:00
bors[bot]	8957251eed	Merge #751 751: Update version for the next release (v0.38.0) in Cargo.toml files r=curquiza a=meili-bot ⚠️ This PR is automatically generated. Check the new version is the expected one before merging. Co-authored-by: curquiza <curquiza@users.noreply.github.com>	2022-12-19 17:02:39 +00:00
curquiza	c72535531b	Update version for the next release (v0.38.0) in Cargo.toml files	2022-12-19 16:35:38 +00:00
bors[bot]	19ee9a828f	Merge #3262 3262: Clippy fixes after updating Rust to v1.66 r=curquiza a=dureuill Ran `cargo clippy --fix` Fixes the CI. Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2022-12-19 14:05:59 +00:00
Louis Dureuil	869d331680	Clippy fixes after updating Rust to v1.66	2022-12-19 14:17:12 +01:00
curquiza	913eff5b2f	Use ubuntu-18.04 container in rust tests	2022-12-19 10:46:29 +01:00
Louis Dureuil	916c23e7be	Tests: rename snapshots	2022-12-19 10:07:17 +01:00
Louis Dureuil	ad9937c755	Fix tests after adding DeletionStrategy	2022-12-19 10:07:17 +01:00
Louis Dureuil	171c942282	Soft-deletion computation no longer takes into account the mapsize Implemented solution 2.3 from https://github.com/meilisearch/meilisearch/issues/3231#issuecomment-1348628824	2022-12-19 10:07:17 +01:00
Louis Dureuil	e2ae3b24aa	Hard or soft delete according to the deletion strategy	2022-12-19 10:00:13 +01:00
Louis Dureuil	fc7618d49b	Add DeletionStrategy	2022-12-19 09:49:58 +01:00
amab8901	b4a73f2d74	Remove redundant date-setting	2022-12-16 08:32:44 +01:00
amab8901	4e175ae882	Replace Index::new_with_creation_dates(...) with Index::new(...)	2022-12-16 08:20:13 +01:00
amab8901	5a0a0468df	Combine created and added into date	2022-12-16 08:11:12 +01:00
ManyTheFish	7f88c4ff2f	Fix #1714 test	2022-12-15 18:22:28 +01:00
ManyTheFish	96d4242b93	Update charabia	2022-12-15 18:22:22 +01:00
ManyTheFish	60ebf0ea0b	Add a specific test on finite pagination placeolder search with distinct attributes	2022-12-15 17:28:20 +01:00
bors[bot]	867279f2a4	Merge #3249 3249: Bring back changes from release-v0.30.3 to main r=curquiza a=curquiza ⚠️ ⚠️ I had to fix git conflicts, ensure I did not lose anything ⚠️ ⚠️ Co-authored-by: Kerollmops <clement@meilisearch.com> Co-authored-by: Tamo <tamo@meilisearch.com> Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2022-12-15 14:13:30 +00:00
bors[bot]	5114686394	Merge #743 743: Fix finite pagination with placeholder search r=Kerollmops a=ManyTheFish this bug is reproducible on real datasets and is hard to isolate in a simple test. related to: https://github.com/meilisearch/meilisearch/issues/3200 poke `@curquiza` Co-authored-by: ManyTheFish <many@meilisearch.com>	2022-12-15 09:31:47 +00:00
ManyTheFish	3322018c06	Fix placeholder search	2022-12-14 20:09:47 +01:00
Louis Dureuil	ce84a59873	Re-apply some changes from #3132	2022-12-14 20:02:39 +01:00
Tamo	d66bb3a53f	rename the two new functions	2022-12-14 17:27:43 +01:00
Tamo	6c0b8edab5	Fix typos Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2022-12-14 17:27:37 +01:00
Tamo	fbbc6eaeca	Fix the import of dumps and snapshot. Some flags were badly applied + the database wrongly deleted when they shouldn't	2022-12-14 17:27:28 +01:00
Kerollmops	60c3bac108	Bump milli to v0.37.3	2022-12-14 17:25:40 +01:00
bors[bot]	9491fe0704	Merge #3247 3247: Re-add push in docker CI r=curquiza a=curquiza I made a mistake here https://github.com/meilisearch/meilisearch/pull/3229, `push` is not `true` by default, see https://github.com/docker/build-push-action#customizing Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com>	2022-12-14 13:15:41 +00:00
Clémentine Urquizar - curqui	240c73d292	Re-add push	2022-12-14 14:05:25 +01:00
amab8901	d3eb8d2d5c	Enable create_raw_index(...) to specify time	2022-12-14 10:44:25 +01:00
bors[bot]	0276d5212a	Merge #728 728: Add some integration tests on the sort criterion r=ManyTheFish a=loiclec This is simply an integration test ensuring that the sort criterion works properly. However, only one version of the algorithm is tested here (the iterative one). To test the version that uses the facet DB, one has to manually set the `CANDIDATES_THRESHOLD` constant to `0`. I have done that and ensured that the test still succeeds. However, in the future, we will probably want to have an option to force which algorithm is used at runtime, for testing purposes. Co-authored-by: Loïc Lecrenier <loic.lecrenier@me.com>	2022-12-14 09:27:12 +00:00
bors[bot]	660be071b5	Merge #3236 3236: Improves clarity of the code that receives payloads r=Kerollmops a=Kerollmops This PR makes small changes to #3164. It improves the clarity and simplicity of some parts of the code. Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-12-13 18:20:24 +00:00
bors[bot]	89542d7d8b	Merge #3241 3241: Remove core mention r=curquiza a=curquiza No impact for the users or the team Co-authored-by: curquiza <clementine@meilisearch.com>	2022-12-13 17:35:50 +00:00
curquiza	f62e7a3501	Remove core mention	2022-12-13 17:34:43 +01:00
Kerollmops	a08cc82983	Revert "Simplify the code when array_each failed" This reverts commit `271685cceb`.	2022-12-13 16:29:49 +01:00
bors[bot]	e2ffc3d69a	Merge #741 741: Add test reproducing the bug fixed by #737 r=Kerollmops a=ManyTheFish related to #737 Co-authored-by: ManyTheFish <many@meilisearch.com>	2022-12-13 15:02:19 +00:00
ManyTheFish	739da9fd4d	Add test	2022-12-13 15:54:43 +01:00
bors[bot]	2af93966e0	Merge #740 740: Fix two nightly errors r=Kerollmops a=irevoire Currently, we have these two errors on rust nightly. It would be nice to help rustc understand what's going on ``` error[E0658]: anonymous lifetimes in `impl Trait` are unstable --> filter-parser/src/lib.rs:173:53 \| 173 \| fn ws<'a, O>(inner: impl FnMut(Span<'a>) -> IResult<O>) -> impl FnMut(Span<'a>) -> IResult<O> { \| ^ expected named lifetime parameter \| = help: add `#![feature(anonymous_lifetime_in_impl_trait)]` to the crate attributes to enable help: consider introducing a named lifetime parameter \| 173 \| fn ws<'a, 'a, O>(inner: impl FnMut(Span<'a>) -> IResult<'a, O>) -> impl FnMut(Span<'a>) -> IResult<O> { \| +++ +++ error[E0658]: anonymous lifetimes in `impl Trait` are unstable --> filter-parser/src/error.rs:36:49 \| 36 \| mut parser: impl FnMut(Span<'a>) -> IResult<O>, \| ^ expected named lifetime parameter \| = help: add `#![feature(anonymous_lifetime_in_impl_trait)]` to the crate attributes to enable help: consider introducing a named lifetime parameter \| 35 ~ pub fn cut_with_err<'a, 'a, O>( 36 ~ mut parser: impl FnMut(Span<'a>) -> IResult<'a, O>, \| For more information about this error, try `rustc --explain E0658`. error: could not compile `filter-parser` due to 2 previous errors ``` Co-authored-by: Tamo <tamo@meilisearch.com>	2022-12-13 14:33:40 +00:00
Kerollmops	7b2f2a4f9c	Do only one convertion to u64	2022-12-13 15:31:55 +01:00
Tamo	2c47500bc3	fix two nightly errors	2022-12-13 15:29:52 +01:00
Kerollmops	5d5615ef45	Rename the ReceivePayload error variant	2022-12-13 15:07:35 +01:00
Kerollmops	526793b5b2	Handle empty arrays the same way we handle other arrays	2022-12-13 14:58:40 +01:00
Kerollmops	271685cceb	Simplify the code when array_each failed	2022-12-13 14:58:05 +01:00
bors[bot]	1af590d3bc	Merge #3234 3234: Update README.md r=curquiza a=tpayet Change Slack link to Discord link Co-authored-by: Thomas Payet <thomas@meilisearch.com>	2022-12-13 11:41:10 +00:00
bors[bot]	dab2634ca8	Merge #3164 3164: Improve the way we receive the documents payload r=Kerollmops a=jiangbo212 # Pull Request ## Related issue Fixes #3037 ## What does this PR do? - writing the playload to a temporary file via BufWritter - deserialising the json tempporary file to an array of Objects by means of a memory map - deserialising thie csv tempporary file by means of a memory map - Adapted some read_json tests ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: jiangbo212 <peiyaoliukuan@gmail.com> Co-authored-by: jiangbo212 <peiyaoliukuan@126.com>	2022-12-13 10:58:24 +00:00
bors[bot]	406ee31d1a	Merge #737 737: Fix typo initial candidates computation r=Kerollmops a=ManyTheFish When `Typo` criterion was after a different criterion than `Words` and the previous criterion wasn't returning any candidates at the first iteration of the bucket sort, then the `initial_candidates` were lost. Now, `Typo`ensure to keep the `initial_candidates` between iterations. related to https://github.com/meilisearch/meilisearch/issues/3200#issuecomment-1345179578 related to https://github.com/meilisearch/meilisearch/issues/3228 Co-authored-by: ManyTheFish <many@meilisearch.com>	2022-12-13 10:29:28 +00:00
ManyTheFish	2d8d0af1a6	Rename short name bc by ic for initial_candidates	2022-12-13 10:56:38 +01:00
Thomas Payet	8a7f90250c	Update README.md Change Slack link to Discord link	2022-12-13 10:46:05 +01:00
bors[bot]	e0a8f8cb5a	Merge #734 734: Fix bug 2945/3021 (missing key in documents database) r=Kerollmops a=loiclec # Pull Request ## Related issue Fixes (partially, until merged into meilisearch) https://github.com/meilisearch/meilisearch/issues/2945 (until we integrate the new milli bump into meilisearch). Note that a dump will not be sufficient to upgrade from meilisearch v0.30.2 to meilisearch v0.30.3 due to this fix because the bug could have caused the `documents` database to be corrupted. Instead, a full manual reimport of the documents will be necessary. ## What does this PR do? There was a bug happening when: 1. A few documents are added to the index 2. Some of these documents are soft-deleted 3. New documents are added, replacing existing ones and triggering a hard-deletion The `IndexDocuments::execute` method would then perform the hard-deletion but forget to change the `external_document_ids` structure appropriately. As a result, the `external_document_ids` would contain keys corresponding to documents that do no exist anymore. To fix this bug, I split the `DeleteDocuments::execute` method into two: `execute_inner` and `execute`. - `execute_inner` returns a `DetailedDocumentDeletionResult` which says whether soft-deletion was used or not - `execute` keeps the exact same signature and behaviour Then, when deleting replaced documents inside `IndexDocuments::execute`, we call `DeleteDocuments::execute_inner` instead of `DeleteDocuments::execute`. If soft-deletion was used, nothing more is done. But if hard-deletion was used, we remove every reference to soft-deleted documents in the new `external_documents_ids` structure. ## Correctness - Every other test still passes - The reproduction test case now passes - In a different branch ([`update-fuzz-test`](https://github.com/meilisearch/milli/pull/735)), I created a fuzz-test that reproduces the past two bugs. This fuzz test cannot find this bug through any combination of some hand-selected `DocumentAddition / DocumentDeletion / DocumentClear / SettingsUpdate` operations. In that test, each relevant operations can be executed with or without soft-deletion, and document additions can be done in batches, replacing or updating existing documents. Co-authored-by: Loïc Lecrenier <loic.lecrenier@me.com>	2022-12-13 09:45:57 +00:00
Loïc Lecrenier	be3b00350c	Apply review suggestions: naming and documentation	2022-12-13 10:15:22 +01:00
jiangbo212	23c1b223b3	Merge branch 'fix-3037' of github.com:jiangbo212/meilisearch into fix-3037	2022-12-13 10:41:50 +08:00
jiangbo212	87ae0032bf	review change	2022-12-13 10:41:43 +08:00
jiangbo212	7c24fea9f2	Merge branch 'main' into fix-3037	2022-12-13 05:16:03 +08:00
ManyTheFish	80d34a4169	Fix typo initial candiddates computation	2022-12-12 19:02:48 +01:00
jiangbo212	27d1bee0bb	Merge branch 'main' into fix-3037-new	2022-12-12 22:16:22 +08:00
jiangbo212	b1c3174061	fix fmt	2022-12-12 22:06:24 +08:00
jiangbo212	fa46dfb7bb	fmt fix	2022-12-12 22:02:56 +08:00
bors[bot]	40d9b73aaf	Merge #3223 3223: Bring back release-v0.30.2 changes into main r=irevoire a=curquiza Only bring back the necessary changes from `release-v0.30.2` to `main`, following v0.30.2 release Co-authored-by: Tamo <tamo@meilisearch.com> Co-authored-by: curquiza <clementine@meilisearch.com>	2022-12-12 13:49:01 +00:00
jiangbo212	169682d3ec	Merge branch 'main' into fix-3037-new	2022-12-12 21:36:10 +08:00
bors[bot]	21b926cb00	Merge #3224 3224: Fix update-cargo-toml-version.yml r=curquiza a=mohitsaxenaknoldus # Pull Request ## Related issue Fixes #3219 ## What does this PR do? - ... ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Mohit Saxena <76725454+mohitsaxenaknoldus@users.noreply.github.com>	2022-12-12 13:27:46 +00:00
Loïc Lecrenier	e3ee553dcc	Remove soft deleted ids from ExternalDocumentIds during document import If the document import replaces a document using hard deletion	2022-12-12 14:16:09 +01:00
bors[bot]	34a6f2598b	Merge #3229 3229: Add a nightly CI: create every day a `nightly` Docker tag based on the latest commit on `main` r=Kerollmops a=curquiza Also, fixes #3195 Easy to follow with the commits - In the Docker CI: - create every day a `nightly` Docker tag based on the latest commit on `main` - check if the release is the latest one, before creating the `latest` Docker tag. A script has been added. - add the `worflow_dispatch` event to trigger the CI to build the `nightly` tag when we want (always on the latest commit on `main`) - In multiple CIs: replace the `released` type by `published`, see [here](https://stackoverflow.com/questions/59319281/github-action-different-between-release-created-and-published) why. Will not impact anything, but will prevent to fail our future automation - Remove a useless CI (code coverage, not used for 1 year) - Remove useless lines (comments and CI logic) that don't have any impact Co-authored-by: curquiza <clementine@meilisearch.com>	2022-12-12 10:46:33 +00:00
curquiza	14824cee86	Remove obsolete comment line	2022-12-11 21:46:48 +01:00
curquiza	796e61ec7e	Remove useless CI	2022-12-11 21:29:23 +01:00
curquiza	9a3f9577b8	Remove useless line in CI	2022-12-11 21:26:05 +01:00
curquiza	2c8eb92537	Check before publish latest	2022-12-11 21:24:52 +01:00
Mohit Saxena	1bf5c0edb9	Update update-cargo-toml-version.yml	2022-12-10 23:04:26 +05:30
curquiza	b1ffbe561e	Add nightly for docker CI	2022-12-09 20:06:59 +01:00
curquiza	84204b8cd5	Replace the released type by published	2022-12-09 19:27:58 +01:00
Mohit Saxena	346fca5608	Update update-cargo-toml-version.yml	2022-12-09 00:20:51 +05:30
Loïc Lecrenier	bebd050961	Add new test for bug 3021	2022-12-08 19:19:40 +01:00
curquiza	4631f4d97f	Bump milli to v0.37.2	2022-12-08 18:16:48 +01:00
Tamo	6f1c30b247	Fix the instance-uid in the data.ms We were writing the instance-uid as bytes instead of string in the data.ms and thus we were unable to parse it later. Also it was less practical for our user to retrieve it and send it to us.	2022-12-08 18:16:43 +01:00
bors[bot]	abba54e913	Merge #3112 3112: Rename meilisearch-http r=Kerollmops a=colbsmcdolbs # Pull Request ## Related issue Fixes #3073 ## What does this PR do? - Renames all references of `meilisearch-http` to `meilisearch` - Might need to be rebased before the 1.0.0 release ## PR checklist Please check if your PR fulfills the following requirements: - [X] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [X] Have you read the contributing guidelines? - [X] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Colby Allen <colbyjayallen@gmail.com>	2022-12-08 16:32:08 +00:00
bors[bot]	c426fa1478	Merge #3212 3212: Setup COMMIT_SHA and COMMIT_DATE build args in the Docker image r=curquiza a=brunoocasali GitHub auto-closed my PR when I synced changes with my remote 🤷‍♂️ https://github.com/meilisearch/meilisearch/pull/2550 The last PR #3205 were closed to help `@curquiza` test the CI. In any case, the summary of changes is quite similar: - Fix `git` usage from my last attempt (when you use `actions/checkout`) you get the `git` command to use. - Add the `build-args` definition from https://github.com/docker/build-push-action#inputs, which is supposed to work precisely as docker build `--build-arg`. Fixes https://github.com/meilisearch/meilisearch/issues/2028 The result will be like this: <img width="556" alt="image" src="https://user-images.githubusercontent.com/4116980/206019608-2713559a-1f58-4ff3-9fec-7720783993ac.png"> Co-authored-by: Bruno Casali <brunoocasali@gmail.com>	2022-12-08 16:07:50 +00:00
bors[bot]	574be942cd	Merge #3221 3221: Update README to reference Meilisearch Cloud r=curquiza a=davelarkan # Pull Request ## Related issue Fixes #3220 ## What does this PR do? - Updates the README to link to the Pricing page where people can choose a Meilisearch Cloud plan ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Dave Larkan <davelarkan@gmail.com>	2022-12-08 15:43:43 +00:00
Colby Allen	2262766494	chore: run fmt nightly on project	2022-12-08 08:31:15 -07:00
Colby Allen	ad2b1467da	Renames meilisearch-http to meilisearch	2022-12-08 08:22:53 -07:00
Dave Larkan	ee37d5e724	Update README to reference Meilisearch Cloud	2022-12-08 15:02:34 +00:00
bors[bot]	ded2a50d14	Merge #3216 3216: Update version for the next release (v1.0.0) in Cargo.toml files r=curquiza a=meili-bot ⚠️ This PR is automatically generated. Check the new version is the expected one before merging. Co-authored-by: curquiza <curquiza@users.noreply.github.com>	2022-12-08 13:49:50 +00:00
Bruno Casali	58327979f1	Use correct env vars "VERGEN_*" on Dockerfile	2022-12-08 10:48:16 -03:00
Bruno Casali	50d9fe036e	Setup COMMIT_SHA and COMMIT_DATE build args in the Docker image	2022-12-08 10:48:16 -03:00
curquiza	026cf223b3	Update version for the next release (v1.0.0) in Cargo.toml files	2022-12-08 12:20:17 +00:00
bors[bot]	af6f7f8462	Merge #3215 3215: Use nightly in cargo fmt r=curquiza a=curquiza Discussed with `@Kerollmops,` needs this change Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com>	2022-12-08 10:53:24 +00:00
Clémentine Urquizar - curqui	5023d36ee7	Use nightly in cargo fmt	2022-12-08 11:51:13 +01:00
bors[bot]	1f1beae077	Merge #729 729: Fix distincted exhaustive hits r=Kerollmops a=ManyTheFish This PR changes the name and behavior of `bucket_candidates`: - `bucket_candidates` become `initial_candidates` that is less confusing - `initial_candidates` is no more a simple `RoaringBitmap` but an enum allowing us to precise if the candidates are exhaustive or not - this enum ensures that any modification is allowed only if the candidates are not already exhaustive. The bug occurred because `initial_candidates` are modified during the bucket sort allowing the estimation to be more and more precise along the search, and this was an issue when the `initial_candidates` were already exhaustive, now, if candidates are exhaustive, then no modifications are made. Co-authored-by: ManyTheFish <many@meilisearch.com>	2022-12-08 09:26:34 +00:00
ManyTheFish	55724f2412	Introduce an initial candidates set that makes the difference between an exhaustive count and an estimation	2022-12-08 09:41:34 +01:00
ManyTheFish	6d50ea0830	add tests	2022-12-08 08:56:57 +01:00
bors[bot]	f4dc4c5d8d	Merge #3210 3210: Fix `MDB_PAGE_FULL` by bumping LMDB r=Kerollmops a=Kerollmops This PR fixes #3062 by upgrading LMDB to the latest version. The changes were made in https://github.com/meilisearch/lmdb/pull/1 and https://github.com/meilisearch/lmdb-rs/pull/12. As heed directly depends on the latest main commit of https://github.com/meilisearch/lmdb-rs, we can bump the `lmdb-rkv-sys` dependency in the Meilisearch _Cargo.lock_ by doing a: ``` cargo update -p lmdb-rkv-sys ``` Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-12-07 16:21:23 +00:00
Loïc Lecrenier	f37c86e0b2	Add some integration tests on the sort criterion	2022-12-07 15:59:33 +01:00
jiangbo212	717dd36547	Merge branch 'fix-3037' of github.com:jiangbo212/meilisearch into fix-3037	2022-12-07 22:54:16 +08:00
jiangbo212	538030c2da	change NameTempFile to tempfile()	2022-12-07 22:47:32 +08:00
bors[bot]	098c410612	Merge #727 727: Fix bug in filter search r=Kerollmops a=loiclec # Pull Request ## Related issue Fixes (partially, until merged into meilisearch) https://github.com/meilisearch/meilisearch/issues/3178 ## What does this PR do? The most important change is this one: ```rust // in milli/src/search/facet/facet_range_search.rs, line 239 let should_stop = { match self.right { Bound::Included(right) => right < previous_key.left_bound, Bound::Excluded(right) => right <= previous_key.left_bound, Bound::Unbounded => false, } }; ``` where the operations `<` and `<=` between the two branches were switched. This caused (very few) documents to be missing from filter results. The second change is a simplification of the algorithm for filters such as `field = value`, where we now perform a direct query into the "Level 0" of the facet db to retrieve the docids instead of invoking the full facet search algorithm. This change is done in `milli/src/search/facet/filter.rs`. I have added yet more insta-snapshot tests, rechecked the content of the snapshots, and added some integration tests as well. This is purely a fix in the search algorithms. Based on this PR alone, a dump will not be necessary to switch from v0.30.1 (where this bug is present) to v0.30.2 (where this PR is merged). Co-authored-by: Loïc Lecrenier <loic.lecrenier@me.com>	2022-12-07 14:34:59 +00:00
Kerollmops	1d5294d11a	Bump lmdb version	2022-12-07 15:29:56 +01:00
bors[bot]	ee10cb8c87	Merge #726 726: Update the contributing.md r=curquiza a=irevoire Co-authored-by: Tamo <tamo@meilisearch.com>	2022-12-07 13:59:04 +00:00
Loïc Lecrenier	d38cc73630	Add one more filter "integration" test	2022-12-07 14:38:25 +01:00
Loïc Lecrenier	e688581c36	Add tests for facet range search on different field ids	2022-12-07 14:38:21 +01:00
Loïc Lecrenier	4ac8f96342	Simplify implementation of equality condition in filters	2022-12-07 14:38:18 +01:00
Loïc Lecrenier	1c9555566e	Fix bug in facet range search	2022-12-07 14:38:14 +01:00
Loïc Lecrenier	303d740245	Prepare fix within facet range search By creating snapshots and updating the format of the existing snapshots. The next commit will apply the fix, which will show its effects cleanly on the old and new snapshot tests	2022-12-07 14:38:10 +01:00
bors[bot]	34c3e5ec5e	Merge #3208 3208: Stop snapshotting the version of meilisearch in the dump r=Kerollmops a=irevoire It might change, and we don't want to update this test every time we make a new release. Co-authored-by: Tamo <tamo@meilisearch.com>	2022-12-07 12:54:55 +00:00
Tamo	1c3a326199	stop snapshotting the version of meilisearch in the dump It might change and we don't want to update this test everytime we make a new release.	2022-12-07 13:26:02 +01:00
Tamo	250743885d	add a sentence about installing rust-nightly	2022-12-07 12:31:43 +01:00
bors[bot]	34c0f11c26	Merge #3207 3207: Add release check when starting latest CI r=Kerollmops a=curquiza Adding this to have the same kind of check before starting to move the latest tag <img width="737" alt="Capture d’écran 2022-12-07 à 12 18 33" src="https://user-images.githubusercontent.com/20380692/206165868-18a2be7c-78ec-48c9-acb9-d7f60797c2e3.png"> Also, removing an un-unused script Co-authored-by: curquiza <clementine@meilisearch.com>	2022-12-07 11:27:47 +00:00
Tamo	5eecb8489d	Update CONTRIBUTING.md Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com>	2022-12-07 12:23:12 +01:00
Tamo	0e5c3b1f64	Update CONTRIBUTING.md Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com>	2022-12-07 12:23:06 +01:00
curquiza	be300138e4	Add release check when starting latest CI	2022-12-07 12:22:44 +01:00
bors[bot]	2ed6017603	Merge #3204 3204: Bring back v0.30.1 changes to `main` r=curquiza a=curquiza I was not able to just import `release-v0.30.1` to `main`, see: <img width="1371" alt="Capture d’écran 2022-12-06 à 20 03 50" src="https://user-images.githubusercontent.com/20380692/206000844-b39b3063-7da2-475f-b3e4-1791c39a7c2f.png"> So I cherry-picked the commits. ⚠️ ⚠️ ⚠️ I had a git conflict here <img width="730" alt="Capture d’écran 2022-12-06 à 20 09 04" src="https://user-images.githubusercontent.com/20380692/206001007-f56bc28f-c0b1-46a0-bb60-cce4e73b9584.png"> ⚠️ ⚠️ ⚠️ Check out carefully how I fixed it Co-authored-by: curquiza <curquiza@users.noreply.github.com> Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-12-07 11:08:37 +00:00
Kerollmops	c1337f9e08	Update dump snap to new version	2022-12-07 11:48:29 +01:00
bors[bot]	9acac28574	Merge #3128 3128: Bumps cargo_toml version to most up to date r=curquiza a=colbsmcdolbs # Pull Request ## Related issue Fixes #3127 ## What does this PR do? - The README of this repository declares that one package is not up to date. In order to ensure Due Diligence, I have bumped the version number of the package. No test failures running on Windows. ## PR checklist Please check if your PR fulfills the following requirements: - [X] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [X] Have you read the contributing guidelines? - [X] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Colby Allen <colbyjayallen@gmail.com>	2022-12-07 10:31:25 +00:00
jiangbo212	cb1d184904	fmt fix	2022-12-07 17:04:24 +08:00
jiangbo212	2841b09789	Merge branch 'meilisearch:main' into fix-3037	2022-12-07 16:30:21 +08:00
jiangbo212	35f3dd68b6	error change and tokio file use change	2022-12-07 16:20:36 +08:00
Kerollmops	f1de3aa75a	Make the tests use MB to trigger page size issues	2022-12-06 20:10:10 +01:00
Kerollmops	e4e4370a3c	Clamp the databases size to the page size	2022-12-06 20:09:49 +01:00
Kerollmops	24c79b79f9	Bump milli to v0.37.1	2022-12-06 20:05:52 +01:00
curquiza	5db7c4057c	Update version for the next release (v0.30.1) in Cargo.toml files	2022-12-06 20:05:46 +01:00
Tamo	f53bdc4320	update the contributing.md	2022-12-06 17:41:05 +01:00
bors[bot]	0a301b5f88	Merge #723 723: Fix bug in handling of soft deleted documents when updating settings r=Kerollmops a=loiclec # Pull Request ## Related issue Fixes (partially, until merged into meilisearch) https://github.com/meilisearch/meilisearch/issues/3021 ## What does this PR do? This PR fixes the bug where a `missing key in documents database` internal error message could appear when indexing documents. When updating the settings, before clearing the database and before creating the transform output, we now modify the `ExternalDocumentsIds` structure to get rid of all references to soft deleted document ids in its FSTs. It used to be that updating the settings would clear the soft-deleted document ids, but keep the original `ExternalDocumentsIds` structure. As a consequence of this, when processing a future document addition, we could wrongly believe that a document was being replaced when, in fact, it was a completely new document. See the tests `bug_3021_first`, `bug_3021_second`, and `bug_3021` for a minimal test case that would have reproduced the issue. We need to take special care to: - evaluate how users should update to v0.30.1 (containing this fix): dump? reimporting all documents from scratch? - understand IF/HOW this bug could have caused duplicate documents to be returned - and evaluate the correctness of the fix, of course :) Co-authored-by: Loïc Lecrenier <loic.lecrenier@me.com>	2022-12-06 14:37:38 +00:00
Loïc Lecrenier	a993b68684	Cargo fmt >:-(	2022-12-06 15:22:10 +01:00
Loïc Lecrenier	80c7a00567	Fix compilation error in tests of settings update	2022-12-06 15:19:26 +01:00
Loïc Lecrenier	67d8cec209	Fix bug in handling of soft deleted documents when updating settings	2022-12-06 15:09:19 +01:00
bors[bot]	2867d2e91a	Merge #3190 3190: Fix the dump date-import of the dumpv4 r=irevoire a=irevoire # Pull Request After merging https://github.com/meilisearch/meilisearch/pull/3012 I realized that the tests on the date of the dump-v4 were still ignored, thus, I fixed them and then noticed #3012 wasn't working properly. ## Related issue Fixes https://github.com/meilisearch/meilisearch/issues/2987 a second time `@funilrys` since you wrote most of the code you might be interested, but don't feel obligated to review this code. Someone from the team will double-check it works 😁 Co-authored-by: Tamo <tamo@meilisearch.com>	2022-12-06 10:47:00 +00:00
bors[bot]	2a846aaae7	Merge #719 719: Add more members of `filter_parser` to `milli::` & `From<&str>` implementation for `Token` r=Kerollmops a=GregoryConrad ## What does this PR do? The current `milli::Filter` and `milli::FilterCondition` APIs require working with some members of `filter_parser` directly that `milli::` does not re-export to its users (at least when not parsing input using `parse`). Also, using `filter_parser` does not make sense when using milli from an embedded context where there is no query to parse. Instead of reworking `milli::Filter` and `milli::FilterCondition`, this PR adds two non-breaking changes that ease the use of milli: - Re-exports more members of the dependent version of `filter_parser` in `milli` - Implements `From<&str>` for `filter_parser::Token` - This will also allow some basic tests that need to create a `Token` from a string to avoid some boilerplate. In conjunction, both of these will allow milli users to easily create a `Token` from a `&str` without needing to add `filter_parser` as an extra dependency. Note: I wanted to use `FromStr` for the `From` implementation; however, it requires returning a `Result` which is not needed for the conversion. Thus, I just left it as `From<&str>`. Co-authored-by: Gregory Conrad <gregorysconrad@gmail.com>	2022-12-06 10:36:00 +00:00
bors[bot]	1458a12531	Merge #3197 3197: Revert "Upgrade alpine 3.16 to 3.17" r=irevoire a=curquiza Reverts meilisearch/meilisearch#3189 Because `rust:alpine3.17` does not exist, and our scheduled CI failed: https://github.com/meilisearch/meilisearch/actions/runs/3626327181 `@ivanionut` for your information, I'm sorry I should have better checked before accepting the PR, this is my bad Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com>	2022-12-06 10:25:11 +00:00
Clémentine Urquizar - curqui	cbb8d0f97b	Revert "Upgrade alpine 3.16 to 3.17"	2022-12-06 11:09:57 +01:00
Tamo	bef81065f9	return the same time in case we didn't found a created or updated at	2022-12-06 11:03:23 +01:00
Tamo	180511795b	Update dump/src/reader/v4/mod.rs fix typo Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2022-12-06 10:53:43 +01:00
bors[bot]	3bef6e6690	Merge #3175 3175: Rename dump command from --dumps-dir to --dump-dir r=dureuill a=dureuill # Pull Request ## Related issue Fixes #3132 ## What does this PR do? - Rename the dump commands, env variables and default config ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2022-12-06 09:49:42 +00:00
bors[bot]	d6eacb2aac	Merge #722 722: Geosearch for zero radius r=irevoire a=amab8901 # Pull Request ## Related issue Fixes #3167 (https://github.com/meilisearch/meilisearch/issues/3167) ## What does this PR do? - allows Geosearch with zero radius to return the specified location when the coordinates match perfectly (instead of returning nothing). See link for more details. - new attempt on https://github.com/meilisearch/milli/pull/713 ## PR checklist Please check if your PR fulfills the following requirements: - [ X ] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [ X ] Have you read the contributing guidelines? - [ X ] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: amab8901 <amab8901@protonmail.com> Co-authored-by: Tamo <irevoire@protonmail.ch>	2022-12-05 19:57:08 +00:00
Tamo	212dbfa3b5	Update milli/src/search/facet/filter.rs	2022-12-05 20:56:21 +01:00
amab8901	456da5de9c	Geosearch for zero radius	2022-12-05 20:11:46 +01:00
bors[bot]	46e26ab550	Merge #720 720: Make soft deletion optional in document addition and deletion + add lots of tests r=irevoire a=loiclec # Pull Request ## What does this PR do? When debugging recent issues, I created a few unit tests in the hopes reproducing the bugs I was looking for. In the end, I didn't find any, but I thought it would still be good to keep those tests. More importantly, I added a field to the `DeleteDocuments` and `IndexDocuments` builders, called `disable_soft_deletion`. If set to `true`, the indexing/deletion will never add documents to the `soft_deleted_documents_ids` and instead perform a real deletion of the documents from the databases. For the new tests, I have: - Improved the insta-snapshot format of the `external_documents_ids` structure - Added more tests for the facet DB indexing, deletion, and search algorithms, making sure to test them when the facet DB contains strings (instead of numbers) as well. - Added more tests for the incremental indexing of the prefix proximity databases. For example, to see if documents are replaced correctly and if common prefixes are deleted correctly. - Added tests that mix soft deletion and hard deletion, including when processing batches of document updates. Co-authored-by: Loïc Lecrenier <loic.lecrenier@me.com>	2022-12-05 18:26:01 +00:00
bors[bot]	9b23885e85	Merge #3188 3188: re-enable the dump test on the dates r=irevoire a=irevoire I just noticed that we have the real date in the dump-v1 contrarily to the dump-v2/3/4/5, thus we can ensure it doesn't change unexpectedly 👍 Co-authored-by: Tamo <tamo@meilisearch.com>	2022-12-05 18:10:22 +00:00
bors[bot]	8b46093117	Merge #3189 3189: Upgrade alpine 3.16 to 3.17 r=curquiza a=ivanionut Upgrade alpine 3.16 to 3.17 Co-authored-by: Ivan Ionut <ivan.ionut@gmail.com>	2022-12-05 17:39:10 +00:00
Tamo	9c89e3dadc	uncomment more test for the dump v4	2022-12-05 18:15:29 +01:00
Tamo	b0cf431614	Fix the dump date-import of the dumpv4	2022-12-05 18:08:35 +01:00
Ivan Ionut	afe520a67e	Upgrade alpine 3.16 to 3.17	2022-12-05 17:49:15 +01:00
Tamo	688911ed34	re-enable the dump test on the dates	2022-12-05 17:05:37 +01:00
bors[bot]	776af129bf	Merge #3012 3012: Extract the dates out of the dumpv4. r=irevoire a=funilrys Hi there, please review this PR that tries to fix #2987. I'm still learning Rust and I found that #2987 is an excellent way for me to read and learn what others do with Rust. So please excuse my semantics ... Stay safe and healthy. --- # Pull Request This patch possibly fixes #2987. This patch introduces a way to fill the IndexMetadata.created_at and IndexMetadata.updated_at keys from the tasks events. This is done by reading the creation date of the first event (created_at) and the creation date of the last event (updated_at). ## Related issue Fixes #2987 ## What does this PR do? - Extract the dates out of the dumpv4. ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: funilrys <contact@funilrys.com>	2022-12-05 15:57:07 +00:00
Louis Dureuil	492fd2829a	use a consistent dump directory name in tests changed from 'dump' to 'dumps' to be consistent with the default settings Co-authored-by: Tamo <tamo@meilisearch.com>	2022-12-05 16:56:28 +01:00
bors[bot]	ffa6d1ed4e	Merge #3186 3186: Update mini-dashboard to v0.2.4 r=curquiza a=mdubus Co-authored-by: Morgane Dubus <30866152+mdubus@users.noreply.github.com>	2022-12-05 14:45:27 +00:00
Morgane Dubus	293efb7485	Update Cargo.toml	2022-12-05 14:54:01 +01:00
Loïc Lecrenier	cda4ba2bb6	Add document import tests	2022-12-05 12:02:49 +01:00
Loïc Lecrenier	ae59d37b75	Improve insta-snap of the external document ids	2022-12-05 10:51:02 +01:00
Loïc Lecrenier	f2cf981641	Add more tests and allow disabling of soft-deletion outside of tests Also allow disabling soft-deletion in the IndexDocumentsConfig	2022-12-05 10:51:01 +01:00
jiangbo212	6766712840	fmt fix	2022-12-04 23:05:34 +08:00
jiangbo212	980776b646	test fail fix	2022-12-04 22:31:23 +08:00
jiangbo212	6bdd37beb8	tokio file write update	2022-12-04 18:25:06 +08:00
Gregory Conrad	50954d31fa	feat: Re-export Span and Token to milli::	2022-12-03 13:37:33 -05:00
Gregory Conrad	1b5b5778c1	feat: Add From<&str> implementation for Token	2022-12-03 13:13:41 -05:00
funilrys	8b6eba4f0b	Apply fmt.	2022-12-03 17:47:02 +01:00
funilrys	e510ace179	fixup! Re-open tasks queue.	2022-12-03 17:41:33 +01:00
funilrys	f056fc118f	Re-open tasks queue. Indeed, before this patch, I was (probably) breaking every usage of the tasks BufReader. This patch solves the issue by reopening the the tasks file every time its needed.	2022-12-03 17:29:41 +01:00
jiangbo212	5a770ffe47	test fail fix	2022-12-03 22:48:38 +08:00
jiangbo212	7b08d700f7	requested changes fix	2022-12-03 18:52:20 +08:00
jiangbo212	c63748723d	Merge branch 'meilisearch:main' into fix-3037	2022-12-03 15:56:41 +08:00
bors[bot]	40339919ad	Merge #3170 #3180 #3181 #3183 3170: Re-enable importing from dumps v1 r=irevoire a=dureuill # Pull Request ## Related issue Fixes #2985 ## What does this PR do? ### User standpoint - Allows importing dumps version 1 (exported with Meilisearch <=v0.20) to modern-day Meilisearch - Tasks of type "Customs" are skipped, with a warning - Tasks of status "enqueued" are skipped, with a warning - The "WordsPosition" ranking rule is skipped when encountered in the ranking rules, with a warning. After an import from a v1 dump, it is recommended that a user checks each index and its settings. ### Implementation standpoint - Add a dump v1 reader based on the one by `@irevoire` - Add a v1_to_v2 compatibility layer based on the v2_to_v3 one - as v2 requires UUIDs, the v1 indexes are mapped to UUIDs built from their position in the metada file: the first index is given UUID all zeroes, the second one UUID `00000000-0000-0000-0000-000000000001`, and so on... This should have no bearing on the final indexes because v6 is not using UUIDs, but this allows us to correctly identify which tasks belong to which index. - Modify the v2_to_v3 compatibility layer to account for the fact that the reader can actually be a v1_to_v2 compat layer - Make some base dump types Clone - impl Display for v2::settings::Criterion ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! 3180: Bump mislav/bump-homebrew-formula-action from 1 to 2 r=curquiza a=dependabot[bot] Bumps [mislav/bump-homebrew-formula-action](https://github.com/mislav/bump-homebrew-formula-action) from 1 to 2. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/mislav/bump-homebrew-formula-action/releases">mislav/bump-homebrew-formula-action's releases</a>.</em></p> <blockquote> <h2>bump-homebrew-formula 2.0</h2> <h2>What's Changed</h2> <ul> <li>Use Node 16 by <a href="https://github.com/chenrui333"><code>`@chenrui333</code></a>` in <a href="https://github-redirect.dependabot.com/mislav/bump-homebrew-formula-action/pull/36">mislav/bump-homebrew-formula-action#36</a></li> <li>Bump minimist from 1.2.5 to 1.2.6 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://github-redirect.dependabot.com/mislav/bump-homebrew-formula-action/pull/33">mislav/bump-homebrew-formula-action#33</a></li> <li>Bump node-fetch from 2.6.6 to 2.6.7 by <a href="https://github.com/dependabot"><code>`@dependabot</code></a>` in <a href="https://github-redirect.dependabot.com/mislav/bump-homebrew-formula-action/pull/34">mislav/bump-homebrew-formula-action#34</a></li> </ul> <h2>New Contributors</h2> <ul> <li><a href="https://github.com/chenrui333"><code>`@chenrui333</code></a>` made their first contribution in <a href="https://github-redirect.dependabot.com/mislav/bump-homebrew-formula-action/pull/36">mislav/bump-homebrew-formula-action#36</a></li> </ul> <h2>bump-homebrew-formula 1.16</h2> <ul> <li>Replaces broken v1.15 tag, thanks <a href="https://github.com/hendrikmaus"><code>`@hendrikmaus</code></a>` <a href="https://github-redirect.dependabot.com/mislav/bump-homebrew-formula-action/issues/32">mislav/bump-homebrew-formula-action#32</a></li> <li>Add <code>push-to</code> option, thanks <a href="https://github.com/codefromthecrypt"><code>`@codefromthecrypt</code></a>` <a href="https://github-redirect.dependabot.com/mislav/bump-homebrew-formula-action/pull/30">mislav/bump-homebrew-formula-action#30</a></li> <li>Fix syntax error, thanks <a href="https://github.com/hendrikmaus"><code>`@hendrikmaus</code></a>` <a href="https://github.com/wata727"><code>`@wata727</code></a>` <a href="https://github-redirect.dependabot.com/mislav/bump-homebrew-formula-action/pull/27">mislav/bump-homebrew-formula-action#27</a></li> <li>Ensure repeated placeholders in <code>commit-message</code> are expanded, thanks <a href="https://github.com/hendrikmaus"><code>`@hendrikmaus</code></a>` <a href="https://github-redirect.dependabot.com/mislav/bump-homebrew-formula-action/pull/29">mislav/bump-homebrew-formula-action#29</a></li> </ul> <h2>bump-homebrew-formula 1.14</h2> <ul> <li>Ignore HTTP 409 error when fast-forwading the main branch of <code>homebrew-tap</code> fork</li> </ul> <h2>bump-homebrew-formula 1.13</h2> <ul> <li>Add <code>create-pullrequest</code> input to control whether or not a PR is submitted to <code>homebrew-tap</code></li> <li>Add <code>download-sha256</code> input to define the SHA256 checksum of the archive at <code>download-url</code></li> <li>Fix creating a new branch in the forked repo failing with HTTP 404</li> </ul> <h2>bump-homebrew-formula 1.12</h2> <ul> <li>Fix Actions CJS loader halting on <code>foo?.bar</code> JS syntax</li> </ul> <h2>bump-homebrew-formula 1.11</h2> <ul> <li>New optional <code>formula-path</code> input accepts the filename of the formula file to edit (default <code>Formula/<formula-name>.rb</code>).</li> <li>Remove <code>revision N</code> lines when bumping Homebrew formulae.</li> </ul> <h2>bump-homebrew-formula 1.10</h2> <ul> <li>The new optional <code>tag-name</code> input allows this action to be <a href="https://docs.github.com/en/actions/managing-workflow-runs/manually-running-a-workflow">manually triggered via <code>workflow_dispatch</code></a> instead of on git push to a tag.</li> </ul> <h2>bump-homebrew-formula 1.9</h2> <ul> <li>Fix following multiple HTTP redirects while calculating checksum for <code>download-url</code></li> </ul> <h2>bump-homebrew-formula 1.8</h2> <ul> <li>Enable JavaScript source maps for better failure debugging</li> </ul> <h2>bump-homebrew-formula 1.7</h2> <ul> <li> <p>Allow <code>download-url</code> as input parameter</p> </li> <li> <p>Add support for git-based <code>download-url</code></p> </li> </ul> <h2>bump-homebrew-formula 1.6</h2> <ul> <li>Control the git commit message template being used for updating the formula file via the <code>commit-message</code> action input</li> </ul> <h2>bump-homebrew-formula 1.5</h2> <ul> <li>Support detection version from <code>https://github.com/OWNER/REPO/releases/download/TAG/FILE</code> download URLs</li> </ul> <!-- raw HTML omitted --> </blockquote> <p>... (truncated)</p> </details> <details> <summary>Commits</summary> <ul> <li><a href="`fcd7e28e54`"><code>fcd7e28</code></a> lib</li> <li><a href="`33989a8502`"><code>33989a8</code></a> Merge branch 'main' into v2</li> <li><a href="`5983bb6c59`"><code>5983bb6</code></a> Improve extracting complex tag names from URLs</li> <li><a href="`9750a1166b`"><code>9750a11</code></a> lib</li> <li><a href="`64410e9c96`"><code>64410e9</code></a> v2</li> <li><a href="`677d7482a3`"><code>677d748</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/mislav/bump-homebrew-formula-action/issues/36">#36</a> from chenrui333/node-16</li> <li><a href="`f364e76079`"><code>f364e76</code></a> also update some build dependencies</li> <li><a href="`c08fd9bee5`"><code>c08fd9b</code></a> deps: update to use nodev16</li> <li><a href="`280f532e9a`"><code>280f532</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/mislav/bump-homebrew-formula-action/issues/34">#34</a> from mislav/dependabot/npm_and_yarn/node-fetch-2.6.7</li> <li><a href="`5d94a66af3`"><code>5d94a66</code></a> Bump node-fetch from 2.6.6 to 2.6.7</li> <li>Additional commits viewable in <a href="https://github.com/mislav/bump-homebrew-formula-action/compare/v1...v2">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=mislav/bump-homebrew-formula-action&package-manager=github_actions&previous-version=1&new-version=2)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) You can trigger a rebase of this PR by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) </details> 3181: Bump Swatinem/rust-cache from 2.0.0 to 2.2.0 r=curquiza a=dependabot[bot] Bumps [Swatinem/rust-cache](https://github.com/Swatinem/rust-cache) from 2.0.0 to 2.2.0. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/Swatinem/rust-cache/releases">Swatinem/rust-cache's releases</a>.</em></p> <blockquote> <h2>v2.2.0</h2> <ul> <li>Add new <code>save-if</code> option to always restore, but only conditionally save the cache.</li> </ul> <h2>v2.1.0</h2> <ul> <li>Only hash <code>Cargo.{lock,toml}</code> files in the configured workspace directories.</li> </ul> <h2>v2.0.2</h2> <ul> <li>Avoid calling cargo metadata on pre-cleanup.</li> <li>Added <code>prefix-key</code>, <code>cache-directories</code> and <code>cache-targets</code> options.</li> </ul> <h2>v2.0.1</h2> <ul> <li>Primarily just updating dependencies to fix GitHub deprecation notices.</li> </ul> </blockquote> </details> <details> <summary>Changelog</summary> <p><em>Sourced from <a href="https://github.com/Swatinem/rust-cache/blob/master/CHANGELOG.md">Swatinem/rust-cache's changelog</a>.</em></p> <blockquote> <h2>2.2.0</h2> <ul> <li>Add new <code>save-if</code> option to always restore, but only conditionally save the cache.</li> </ul> <h2>2.1.0</h2> <ul> <li>Only hash <code>Cargo.{lock,toml}</code> files in the configured workspace directories.</li> </ul> <h2>2.0.2</h2> <ul> <li>Avoid calling <code>cargo metadata</code> on pre-cleanup.</li> <li>Added <code>prefix-key</code>, <code>cache-directories</code> and <code>cache-targets</code> options.</li> </ul> <h2>2.0.1</h2> <ul> <li>Primarily just updating dependencies to fix GitHub deprecation notices.</li> </ul> </blockquote> </details> <details> <summary>Commits</summary> <ul> <li><a href="`359a70e43a`"><code>359a70e</code></a> 2.2.0</li> <li><a href="`ecee04e7b3`"><code>ecee04e</code></a> feat: add save-if option, closes <a href="https://github-redirect.dependabot.com/Swatinem/rust-cache/issues/66">#66</a> (<a href="https://github-redirect.dependabot.com/Swatinem/rust-cache/issues/91">#91</a>)</li> <li><a href="`b894d59a8d`"><code>b894d59</code></a> 2.1.0</li> <li><a href="`e78327dd9e`"><code>e78327d</code></a> small code style improvements, README and CHANGELOG updates</li> <li><a href="`ccdddcc049`"><code>ccdddcc</code></a> only hash Cargo.toml/Cargo.lock that belong to a configured workspace (<a href="https://github-redirect.dependabot.com/Swatinem/rust-cache/issues/90">#90</a>)</li> <li><a href="`b5ec9edd91`"><code>b5ec9ed</code></a> 2.0.2</li> <li><a href="`3f2513fdf4`"><code>3f2513f</code></a> avoid calling cargo metadata on pre-cleanup</li> <li><a href="`19c46583c5`"><code>19c4658</code></a> update dependencies</li> <li><a href="`b8e72aae83`"><code>b8e72aa</code></a> Added <code>prefix-key</code> <code>cache-directories</code> and <code>cache-targets</code> options (<a href="https://github-redirect.dependabot.com/Swatinem/rust-cache/issues/85">#85</a>)</li> <li><a href="`22c9328bcb`"><code>22c9328</code></a> 2.0.1</li> <li>Additional commits viewable in <a href="https://github.com/Swatinem/rust-cache/compare/v2.0.0...v2.2.0">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=Swatinem/rust-cache&package-manager=github_actions&previous-version=2.0.0&new-version=2.2.0)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) You can trigger a rebase of this PR by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) </details> 3183: Use ubuntu-latest when not impacting r=Kerollmops a=curquiza Minor changes - Use `ubuntu-latest` for CI where there is no compilation - rename one of the workflow (obsolete name) Co-authored-by: Louis Dureuil <louis@meilisearch.com> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com> Co-authored-by: curquiza <clementine@meilisearch.com>	2022-12-01 17:33:59 +00:00
curquiza	7f3653ec31	Use ubuntu-latest when not impacting	2022-12-01 18:23:29 +01:00
bors[bot]	a1f5ec1e9e	Merge #3179 3179: Bump svenstaro/upload-release-action from 1.pre.release to 2.3.0 r=curquiza a=dependabot[bot] Bumps [svenstaro/upload-release-action](https://github.com/svenstaro/upload-release-action) from 1.pre.release to 2.3.0. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/svenstaro/upload-release-action/releases">svenstaro/upload-release-action's releases</a>.</em></p> <blockquote> <h2>2.3.0</h2> <ul> <li>Now defaults <code>repo_token</code> to <code>${{ github.token }}</code> and <code>tag</code> to <code>${{ github.ref }}</code> <a href="https://github-redirect.dependabot.com/svenstaro/upload-release-action/pull/69">#69</a> (thanks <a href="https://github.com/leighmcculloch"><code>`@leighmcculloch</code></a>)**</li>` </ul> <h2>2.2.1</h2> <ul> <li>Added support for the GitHub pagination API for repositories with many releases <a href="https://github-redirect.dependabot.com/svenstaro/upload-release-action/pull/36">#36</a> (thanks <a href="https://github.com/djpohly"><code>`@djpohly</code></a>)</li>` </ul> <h2>2.2.0</h2> <ul> <li>Add support for ceating a new release in a foreign repository <a href="https://github-redirect.dependabot.com/svenstaro/upload-release-action/pull/25">#25</a> (thanks <a href="https://github.com/kittaakos"><code>`@kittaakos</code></a>)</li>` <li>Upgrade all deps</li> </ul> <h2>2.1.1</h2> <ul> <li>Fix <code>release_name</code> option <a href="https://github-redirect.dependabot.com/svenstaro/upload-release-action/pull/27">#27</a> (thanks <a href="https://github.com/kittaakos"><code>`@kittaakos</code></a>)</li>` </ul> <h2>2.1.0</h2> <ul> <li>Strip refs/heads/ from the input tag <a href="https://github-redirect.dependabot.com/svenstaro/upload-release-action/pull/23">#23</a> (thanks <a href="https://github.com/OmarEmaraDev"><code>`@OmarEmaraDev</code></a>)</li>` </ul> <h2>2.0.0</h2> <ul> <li>Add <code>prerelease</code> input parameter. Setting this marks the created release as a pre-release.</li> <li>Add <code>release_name</code> input parameter. Setting this explicitly sets the title of the release.</li> <li>Add <code>body</code> input parameter. Setting this sets the text content of the created release.</li> <li>Add <code>browser_download_url</code> output variable. This contains the publicly accessible download URL of the uploaded artifact.</li> <li>Allow for leaving <code>asset_name</code> unset. This will cause the asset to use the filename.</li> </ul> <h2>1.1.0</h2> <p>No release notes provided.</p> <h2>1.0.1</h2> <p>No release notes provided.</p> <h2>1.0.0</h2> <p>No release notes provided.</p> </blockquote> </details> <details> <summary>Changelog</summary> <p><em>Sourced from <a href="https://github.com/svenstaro/upload-release-action/blob/master/CHANGELOG.md">svenstaro/upload-release-action's changelog</a>.</em></p> <blockquote> <h2>[2.3.0] - 2022-06-05</h2> <ul> <li>Now defaults <code>repo_token</code> to <code>${{ github.token }}</code> and <code>tag</code> to <code>${{ github.ref }}</code> <a href="https://github-redirect.dependabot.com/svenstaro/upload-release-action/pull/69">#69</a> (thanks <a href="https://github.com/leighmcculloch"><code>`@leighmcculloch</code></a>)</li>` </ul> <h2>[2.2.1] - 2020-12-16</h2> <ul> <li>Added support for the GitHub pagination API for repositories with many releases <a href="https://github-redirect.dependabot.com/svenstaro/upload-release-action/pull/36">#36</a> (thanks <a href="https://github.com/djpohly"><code>`@djpohly</code></a>)</li>` </ul> <h2>[2.2.0] - 2020-10-07</h2> <ul> <li>Add support for ceating a new release in a foreign repository <a href="https://github-redirect.dependabot.com/svenstaro/upload-release-action/pull/25">#25</a> (thanks <a href="https://github.com/kittaakos"><code>`@kittaakos</code></a>)</li>` <li>Upgrade all deps</li> </ul> <h2>[2.1.1] - 2020-09-25</h2> <ul> <li>Fix <code>release_name</code> option <a href="https://github-redirect.dependabot.com/svenstaro/upload-release-action/pull/27">#27</a> (thanks <a href="https://github.com/kittaakos"><code>`@kittaakos</code></a>)</li>` </ul> <h2>[2.1.0] - 2020-08-10</h2> <ul> <li>Strip refs/heads/ from the input tag <a href="https://github-redirect.dependabot.com/svenstaro/upload-release-action/pull/23">#23</a> (thanks <a href="https://github.com/OmarEmaraDev"><code>`@OmarEmaraDev</code></a>)</li>` </ul> <h2>[2.0.0] - 2020-07-03</h2> <ul> <li>Add <code>prerelease</code> input parameter. Setting this marks the created release as a pre-release.</li> <li>Add <code>release_name</code> input parameter. Setting this explicitly sets the title of the release.</li> <li>Add <code>body</code> input parameter. Setting this sets the text content of the created release.</li> <li>Add <code>browser_download_url</code> output variable. This contains the publicly accessible download URL of the uploaded artifact.</li> <li>Allow for leaving <code>asset_name</code> unset. This will cause the asset to use the filename.</li> </ul> </blockquote> </details> <details> <summary>Commits</summary> <ul> <li><a href="`133984371c`"><code>1339843</code></a> 2.3.0</li> <li><a href="`c2b649c57e`"><code>c2b649c</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/svenstaro/upload-release-action/issues/72">#72</a> from svenstaro/dependabot/npm_and_yarn/node-fetch-2.6.7</li> <li><a href="`eb625cd0ad`"><code>eb625cd</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/svenstaro/upload-release-action/issues/71">#71</a> from svenstaro/dependabot/npm_and_yarn/ansi-regex-4.1.1</li> <li><a href="`99cbd251b2`"><code>99cbd25</code></a> Bump node-fetch from 2.6.1 to 2.6.7</li> <li><a href="`a6824c9e54`"><code>a6824c9</code></a> Bump ansi-regex from 4.1.0 to 4.1.1</li> <li><a href="`6eb74c809d`"><code>6eb74c8</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/svenstaro/upload-release-action/issues/67">#67</a> from svenstaro/dependabot/npm_and_yarn/minimist-1.2.6</li> <li><a href="`8ec375d911`"><code>8ec375d</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/svenstaro/upload-release-action/issues/69">#69</a> from leighmcculloch/patch-1</li> <li><a href="`8d45355ac2`"><code>8d45355</code></a> Update README.md</li> <li><a href="`7d269bd712`"><code>7d269bd</code></a> Add defaults to commonly set fields to reduce setup</li> <li><a href="`c71fb95114`"><code>c71fb95</code></a> Bump minimist from 1.2.5 to 1.2.6</li> <li>Additional commits viewable in <a href="https://github.com/svenstaro/upload-release-action/compare/v1-release...2.3.0">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=svenstaro/upload-release-action&package-manager=github_actions&previous-version=1.pre.release&new-version=2.3.0)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) You can trigger a rebase of this PR by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) </details> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2022-12-01 17:11:54 +00:00
dependabot[bot]	dd6593f7b6	Bump Swatinem/rust-cache from 2.0.0 to 2.2.0 Bumps [Swatinem/rust-cache](https://github.com/Swatinem/rust-cache) from 2.0.0 to 2.2.0. - [Release notes](https://github.com/Swatinem/rust-cache/releases) - [Changelog](https://github.com/Swatinem/rust-cache/blob/master/CHANGELOG.md) - [Commits](https://github.com/Swatinem/rust-cache/compare/v2.0.0...v2.2.0) --- updated-dependencies: - dependency-name: Swatinem/rust-cache dependency-type: direct:production update-type: version-update:semver-minor ... Signed-off-by: dependabot[bot] <support@github.com>	2022-12-01 17:02:25 +00:00
dependabot[bot]	fc905f29e9	Bump mislav/bump-homebrew-formula-action from 1 to 2 Bumps [mislav/bump-homebrew-formula-action](https://github.com/mislav/bump-homebrew-formula-action) from 1 to 2. - [Release notes](https://github.com/mislav/bump-homebrew-formula-action/releases) - [Commits](https://github.com/mislav/bump-homebrew-formula-action/compare/v1...v2) --- updated-dependencies: - dependency-name: mislav/bump-homebrew-formula-action dependency-type: direct:production update-type: version-update:semver-major ... Signed-off-by: dependabot[bot] <support@github.com>	2022-12-01 17:02:18 +00:00
dependabot[bot]	f892d122de	Bump svenstaro/upload-release-action from 1.pre.release to 2.3.0 Bumps [svenstaro/upload-release-action](https://github.com/svenstaro/upload-release-action) from 1.pre.release to 2.3.0. - [Release notes](https://github.com/svenstaro/upload-release-action/releases) - [Changelog](https://github.com/svenstaro/upload-release-action/blob/master/CHANGELOG.md) - [Commits](https://github.com/svenstaro/upload-release-action/compare/v1-release...2.3.0) --- updated-dependencies: - dependency-name: svenstaro/upload-release-action dependency-type: direct:production ... Signed-off-by: dependabot[bot] <support@github.com>	2022-12-01 17:02:14 +00:00
bors[bot]	7e314664fc	Merge #3169 3169: Improve download-latest.sh script: integrate apple-silicon binary r=curquiza a=curquiza Fixes multiples issues: https://github.com/meilisearch/meilisearch/issues/3044, https://github.com/meilisearch/meilisearch/issues/3027, https://github.com/meilisearch/meilisearch/issues/2613, https://github.com/meilisearch/meilisearch/issues/3014 Improvement/Addition - https://github.com/meilisearch/meilisearch/issues/3044 - https://github.com/meilisearch/meilisearch/issues/3027 Simplification - With https://github.com/meilisearch/meilisearch/issues/3014, we removed the complex part to get the latest version of Meilisearch we can now get directly with the GitHub API Bug fixes: - Should remove the problem the users encountered in this issue (https://github.com/meilisearch/meilisearch/issues/2613) because the related part has been removed from the script Co-authored-by: curquiza <clementine@meilisearch.com> Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com>	2022-12-01 16:18:38 +00:00
Louis Dureuil	a92d67b79f	Also update analytics	2022-12-01 16:24:02 +01:00
Louis Dureuil	daa0f2e9aa	Change dumps_dir to dump_dir in config.toml	2022-12-01 15:46:54 +01:00
Louis Dureuil	e35db5e59b	Rename `dumps-dir` to `dump-dir` in CLI Also rename the associated environment variable	2022-12-01 15:46:54 +01:00
bors[bot]	d3731dda48	Merge #706 706: Limit the reindexing caused by updating settings when not needed r=curquiza a=GregoryConrad ## What does this PR do? When updating index settings using `update::Settings`, sometimes a `reindex` of `update::Settings` is triggered when it doesn't need to be. This PR aims to prevent those unnecessary `reindex` calls. For reference, here is a snippet from the current `execute` method in `update::Settings`: ```rust // ... if stop_words_updated \|\| faceted_updated \|\| synonyms_updated \|\| searchable_updated \|\| exact_attributes_updated { self.reindex(&progress_callback, &should_abort, old_fields_ids_map)?; } ``` - [x] `faceted_updated` - looks good as-is ✅ - [x] `stop_words_updated` - looks good as-is ✅ - [x] `synonyms_updated` - looks good as-is ✅ - [x] `searchable_updated` - fixed in this PR - [x] `exact_attributes_updated` - fixed in this PR ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Gregory Conrad <gregorysconrad@gmail.com>	2022-12-01 13:58:02 +00:00
Louis Dureuil	b727fe7179	Fix integration tests	2022-12-01 14:03:15 +01:00
Louis Dureuil	1260c32d18	Add reader mod test	2022-12-01 14:03:15 +01:00
Louis Dureuil	e03c216952	Add compat_v1_to_v2 test	2022-12-01 14:03:15 +01:00
Louis Dureuil	c7749127fa	Use reader v1 and compat to v2	2022-12-01 14:03:15 +01:00
Louis Dureuil	c8841344e2	v1: RankingRule::from_str	2022-12-01 14:03:15 +01:00
Louis Dureuil	b8de369e33	Add v1 reader	2022-12-01 14:03:15 +01:00
bors[bot]	93b59d55e3	Merge #3172 3172: Add CI to push a latest git tag for every stable Meilisearch release r=curquiza a=curquiza Fixes partially #3147. - Add a CI to add/update a `latest` git tag ONLY when releasing a stable version of Meilisearch (not for pre-release or for custom tag for prototypes) - Update the Docker CI to avoid being triggered when `latest` git tag is pushed: the CI is still triggered when a git tag is pushed, except for the `latest` tag. Reminder: the `latest` Docker image is already created and pushed when releasing a stable version of Meilisearch. This step is already present in our current Docker CI. Co-authored-by: curquiza <clementine@meilisearch.com>	2022-12-01 12:37:09 +00:00
curquiza	b9a8533de1	Add CI to push a latest git tag for every stable Meilisearch release	2022-12-01 12:53:44 +01:00
bors[bot]	51a2613c5c	Merge #715 715: Fix benchmark CI r=irevoire a=curquiza Fixes #714 Tested with our actions: https://github.com/meilisearch/milli/actions/runs/3591527753/jobs/6046157141 Co-authored-by: curquiza <clementine@meilisearch.com>	2022-12-01 10:39:38 +00:00
Louis Dureuil	d44652209d	impl Display for v2::settings::Criterion	2022-12-01 11:15:57 +01:00
Louis Dureuil	5d22c7bcce	Make some dump types Clone	2022-12-01 11:15:57 +01:00
bors[bot]	82e1c4f468	Merge #716 716: Bump Swatinem/rust-cache from 2.0.1 to 2.2.0 r=curquiza a=dependabot[bot] Bumps [Swatinem/rust-cache](https://github.com/Swatinem/rust-cache) from 2.0.1 to 2.2.0. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/Swatinem/rust-cache/releases">Swatinem/rust-cache's releases</a>.</em></p> <blockquote> <h2>v2.2.0</h2> <ul> <li>Add new <code>save-if</code> option to always restore, but only conditionally save the cache.</li> </ul> <h2>v2.1.0</h2> <ul> <li>Only hash <code>Cargo.{lock,toml}</code> files in the configured workspace directories.</li> </ul> <h2>v2.0.2</h2> <ul> <li>Avoid calling cargo metadata on pre-cleanup.</li> <li>Added <code>prefix-key</code>, <code>cache-directories</code> and <code>cache-targets</code> options.</li> </ul> </blockquote> </details> <details> <summary>Changelog</summary> <p><em>Sourced from <a href="https://github.com/Swatinem/rust-cache/blob/master/CHANGELOG.md">Swatinem/rust-cache's changelog</a>.</em></p> <blockquote> <h2>2.2.0</h2> <ul> <li>Add new <code>save-if</code> option to always restore, but only conditionally save the cache.</li> </ul> <h2>2.1.0</h2> <ul> <li>Only hash <code>Cargo.{lock,toml}</code> files in the configured workspace directories.</li> </ul> <h2>2.0.2</h2> <ul> <li>Avoid calling <code>cargo metadata</code> on pre-cleanup.</li> <li>Added <code>prefix-key</code>, <code>cache-directories</code> and <code>cache-targets</code> options.</li> </ul> <h2>2.0.1</h2> <ul> <li>Primarily just updating dependencies to fix GitHub deprecation notices.</li> </ul> <h2>2.0.0</h2> <ul> <li>The action code was refactored to allow for caching multiple workspaces and different <code>target</code> directory layouts.</li> <li>The <code>working-directory</code> and <code>target-dir</code> input options were replaced by a single <code>workspaces</code> option that has the form of <code>$workspace -> $target</code>.</li> <li>Support for considering <code>env-vars</code> as part of the cache key.</li> <li>The <code>sharedKey</code> input option was renamed to <code>shared-key</code> for consistency.</li> </ul> <h2>1.4.0</h2> <ul> <li>Clean both <code>debug</code> and <code>release</code> target directories.</li> </ul> <h2>1.3.0</h2> <ul> <li>Use Rust toolchain file as additional cache key.</li> <li>Allow for a configurable target-dir.</li> </ul> <h2>1.2.0</h2> <ul> <li>Cache <code>~/.cargo/bin</code>.</li> <li>Support for custom <code>$CARGO_HOME</code>.</li> <li>Add a <code>cache-hit</code> output.</li> <li>Add a new <code>sharedKey</code> option that overrides the automatic job-name based key.</li> </ul> <h2>1.1.0</h2> <ul> <li>Add a new <code>working-directory</code> input.</li> <li>Support caching git dependencies.</li> <li>Lots of other improvements.</li> </ul> <h2>1.0.2</h2> <!-- raw HTML omitted --> </blockquote> <p>... (truncated)</p> </details> <details> <summary>Commits</summary> <ul> <li><a href="`359a70e43a`"><code>359a70e</code></a> 2.2.0</li> <li><a href="`ecee04e7b3`"><code>ecee04e</code></a> feat: add save-if option, closes <a href="https://github-redirect.dependabot.com/Swatinem/rust-cache/issues/66">#66</a> (<a href="https://github-redirect.dependabot.com/Swatinem/rust-cache/issues/91">#91</a>)</li> <li><a href="`b894d59a8d`"><code>b894d59</code></a> 2.1.0</li> <li><a href="`e78327dd9e`"><code>e78327d</code></a> small code style improvements, README and CHANGELOG updates</li> <li><a href="`ccdddcc049`"><code>ccdddcc</code></a> only hash Cargo.toml/Cargo.lock that belong to a configured workspace (<a href="https://github-redirect.dependabot.com/Swatinem/rust-cache/issues/90">#90</a>)</li> <li><a href="`b5ec9edd91`"><code>b5ec9ed</code></a> 2.0.2</li> <li><a href="`3f2513fdf4`"><code>3f2513f</code></a> avoid calling cargo metadata on pre-cleanup</li> <li><a href="`19c46583c5`"><code>19c4658</code></a> update dependencies</li> <li><a href="`b8e72aae83`"><code>b8e72aa</code></a> Added <code>prefix-key</code> <code>cache-directories</code> and <code>cache-targets</code> options (<a href="https://github-redirect.dependabot.com/Swatinem/rust-cache/issues/85">#85</a>)</li> <li>See full diff in <a href="https://github.com/Swatinem/rust-cache/compare/v2.0.1...v2.2.0">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=Swatinem/rust-cache&package-manager=github_actions&previous-version=2.0.1&new-version=2.2.0)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) You can trigger a rebase of this PR by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) </details> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2022-12-01 10:08:58 +00:00
curquiza	5bdf5c0aaf	Update the steps to set variables	2022-12-01 11:07:54 +01:00
dependabot[bot]	282b2e3b98	Bump Swatinem/rust-cache from 2.0.1 to 2.2.0 Bumps [Swatinem/rust-cache](https://github.com/Swatinem/rust-cache) from 2.0.1 to 2.2.0. - [Release notes](https://github.com/Swatinem/rust-cache/releases) - [Changelog](https://github.com/Swatinem/rust-cache/blob/master/CHANGELOG.md) - [Commits](https://github.com/Swatinem/rust-cache/compare/v2.0.1...v2.2.0) --- updated-dependencies: - dependency-name: Swatinem/rust-cache dependency-type: direct:production update-type: version-update:semver-minor ... Signed-off-by: dependabot[bot] <support@github.com>	2022-12-01 10:02:54 +00:00
bors[bot]	5e754b3ee0	Merge #708 708: Reduce memory usage of the MatchingWords structure r=ManyTheFish a=loiclec # Pull Request ## Related issue Fixes (partially) https://github.com/meilisearch/meilisearch/issues/3115 ## What does this PR do? 1. Reduces the memory usage caused by the creation of a 10-word query tree by 20x. This is done by deduplicating the `MatchingWord` values, which are heavy because of their inner DFA. The deduplication works by wrapping each `MatchingWord` in a reference-counted box and using a hash map to determine whether a `MatchingWord` DFA already exists for a certain signature, or whether a new one needs to be built. 2. Avoid the worst-case scenario of creating a `MatchingWord` for extremely long words that cannot be indexed by milli. Co-authored-by: Loïc Lecrenier <loic.lecrenier@me.com>	2022-11-30 17:47:34 +00:00
bors[bot]	e1612fcb01	Merge #712 712: Fix bulk facet indexing bug r=Kerollmops a=loiclec # Pull Request ## Related issue Fixes (partially, until merged into meilisearch) https://github.com/meilisearch/meilisearch/issues/3165 ## What does this PR do? Fixes a bug where indexing certain numbers of filterable attribute values in bulk led to corrupted facet databases. This was due to a lossy integer conversion which would ultimately prevent entire levels of the facet database to be written into LMDB. More specifically, this change was made: ```diff - if cur_writer_len as u8 >= self.min_level_size { + if cur_writer_len >= self.min_level_size as usize { ``` I also checked other comparisons to `min_level_size` and other conversions such as `x as u8` in this part of the codebase. Co-authored-by: Loïc Lecrenier <loic.lecrenier@me.com>	2022-11-30 16:51:48 +00:00
Clémentine Urquizar - curqui	da8044f91e	Update download-latest.sh Co-authored-by: Louis Dureuil <louis.dureuil@gmail.com>	2022-11-30 16:55:32 +01:00
Clémentine Urquizar - curqui	9e3b1eb7a8	Update download-latest.sh Co-authored-by: Louis Dureuil <louis.dureuil@gmail.com>	2022-11-30 16:55:27 +01:00
Loïc Lecrenier	9dd4b33a9a	Fix bulk facet indexing bug	2022-11-30 14:27:36 +01:00
curquiza	eab1156f8c	Update script to be used with macOS apple silicon	2022-11-30 14:14:46 +01:00
curquiza	8d405fad12	Simplify download-latest.sh script	2022-11-30 13:53:12 +01:00
jiangbo212	bf96b6df93	clippy fix change	2022-11-30 17:59:06 +08:00
jiangbo212	9c28632498	Merge branch 'main' into fix-3037	2022-11-30 09:38:01 +08:00
jiangbo212	38982d13fe	fix issue 3037	2022-11-30 00:03:22 +08:00
bors[bot]	de22116b3d	Merge #711 711: Replace deprecated gh actions r=curquiza a=pnhatminh # Pull Request ## Related issue Fixes #678 ## What does this PR do? - Replace deprecated github action command with newly defined command. ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Minh Pham <minh.pham@codelink.io>	2022-11-29 09:56:22 +00:00
bors[bot]	6150aa73b0	Merge #3148 3148: Bring back `release-v0.30.0` into `main` r=Kerollmops a=curquiza Following this message https://github.com/meilisearch/meilisearch/pull/3145#issuecomment-1329296168 Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com> Co-authored-by: Tamo <tamo@meilisearch.com> Co-authored-by: bors[bot] <26634292+bors[bot]@users.noreply.github.com>	2022-11-29 09:23:44 +00:00
Minh Pham	5f78522044	Updagte	2022-11-29 10:11:38 +07:00
Gregory Conrad	87e2bc3bed	fix(reindex): reindex in a few more cases Cases: whenever searchable_fields OR user_defined_searchable_fields is modified	2022-11-28 13:12:19 -05:00
bors[bot]	14840a24c6	Merge #3149 3149: Fix the dump tests r=Kerollmops a=irevoire You'll need to trust me on this one. But the tests in the release-v0.30.0 were deactivated for a long time, and I don't know what was wrong with them. Anyway, I checked that these SHA did match the tasks view we're expecting, and it looks good to me. Co-authored-by: Tamo <tamo@meilisearch.com>	2022-11-28 15:48:36 +00:00
Tamo	41eb986f65	Fix the dump tests	2022-11-28 16:39:22 +01:00
Loïc Lecrenier	61b58b115a	Don't create partial matching words for synonyms in ngrams	2022-11-28 16:32:28 +01:00
Clémentine Urquizar - curqui	457a473b72	Bring back `release-v0.30.0` into `release-v0.30.0-temp` (final: into `main`) (#3145 ) * Fix error code of the "duplicate index found" error * Use the content of the ProcessingTasks in the tasks cancelation system * Change the missing_filters error code into missing_task_filters * WIP Introduce the invalid_task_uid error code * Use more precise error codes/message for the task routes + Allow star operator in delete/cancel tasks + rename originalQuery to originalFilters + Display error/canceled_by in task view even when they are = null + Rename task filter fields by using their plural forms + Prepare an error code for canceledBy filter + Only return global tasks if the API key action `index.` is there Add canceledBy task filter * Update tests following task API changes * Rename original_query to original_filters everywhere * Update more insta-snap tests * Make clippy happy They're a happy clip now. * Make rustfmt happy >:-( * Fix Index name parsing error message to fit the specification * Bump milli version to 0.35.1 * Fix the new error messages * fix the error messages and add tests * rename the error codes for the sake of consistency * refactor the way we send the cli informations + add the analytics for the config file and ssl usage * Apply suggestions from code review Co-authored-by: Clément Renault <clement@meilisearch.com> * add a comment over the new infos structure * reformat, sorry @kero * Store analytics for the documents deletions * Add analytics on all the settings * Spawn threads with names * Spawn rayon threads with names * update the distinct attributes to the spec update * update the analytics on the search route * implements the analytics on the health and version routes * Fix task details serialization * Add the question mark to the task deletion query filter * Add the question mark to the task cancelation query filter * Fix tests * add analytics on the task route * Add all the missing fields of the new task query type * Create a new analytics for the task deletion * Create a new analytics for the task creation * batch the tasks seen events * Update the finite pagination analytics * add the analytics of the swap-indexes route * Stop removing the DB when failing to read it * Rename originalFilters into originalFilters * Rename matchedDocuments into providedIds * Add `workflow_dispatch` to flaky.yml * Bump grenad to 0.4.4 * Bump milli to version v0.37.0 * Don't multiply total memory returned by sysinfo anymore sysinfo now returns bytes rather than KB * Add a dispatch to the publish binaries workflow * Fix publish release CI * Don't use gold but the default linker * Always display details for the indexDeletion task * Fix the insta tests * refactorize the whole test suite 1. Make a call to assert_internally_consistent automatically when snapshoting the scheduler. There is no point in snapshoting something broken and expect the dumb humans to notice. 2. Replace every possible call to assert_internally_consistent by a snapshot of the scheduler. It takes as many lines and ensure we never change something without noticing in any tests ever. 3. Name every snapshots: it's easier to debug when something goes wrong and easier to review in general. 4. Stop skipping breakpoints, it's too easy to miss something. Now you must explicitely show which path is the scheduler supposed to use. 5. Add a timeout on the channel.recv, it eases the process of writing tests, now when something file you get a failure instead of a deadlock. * rebase on release-v0.30 * makes clippy happy * update the snapshots after a rebase * try to remove the flakyness of the failing test * Add more analytics on the ranking rules positions * Update the dump test to check for the dumpUid dumpCreation task details * send the ranking rules as a string because amplitude is too dumb to process an array as a single value * Display a null dumpUid until we computed the dump itself on disk * Update tests * Check if the master key is missing before returning an error Co-authored-by: Loïc Lecrenier <loic.lecrenier@me.com> Co-authored-by: bors[bot] <26634292+bors[bot]@users.noreply.github.com> Co-authored-by: Kerollmops <clement@meilisearch.com> Co-authored-by: ManyTheFish <many@meilisearch.com> Co-authored-by: Tamo <tamo@meilisearch.com> Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2022-11-28 16:27:41 +01:00
Gregory Conrad	d3182f3830	refactor: Change return type to keep consistency with others	2022-11-28 10:02:03 -05:00
bors[bot]	f698e6cfdf	Merge #707 707: Add all_obkv_to_json function r=Kerollmops a=GregoryConrad ## What does this PR do? When embedding milli in an application (other than Meilisearch), it often makes sense to not use the `displayed_attributes` functionality and instead just use milli as a full document store. Thus, this PR adds a function, `all_obkv_to_json`, to supplement the already exposed `milli::obkv_to_json` so that those embedding milli do not need to deal with `displayed_attributes` if they don't need to. ~This PR also introduces a slight breaking change: `obkv_to_json` now accepts a reference to `obkv::KvReaderU16` instead of taking ownership of it. As far as I can tell, this seems like a change for the better (`obkv_to_json` only acts upon `obkv` rather than consuming it), but I can change it back if you so desire.~ (reverted in [`935a724`](`935a724c57`)) ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Gregory Conrad <gregorysconrad@gmail.com>	2022-11-28 14:52:45 +00:00
Loïc Lecrenier	f70856bab1	Remove memory usage test that fails when many tests are run in parallel	2022-11-28 12:55:28 +01:00
Loïc Lecrenier	80588daae5	Fix compilation error in formatting benches	2022-11-28 10:27:15 +01:00
Loïc Lecrenier	e2ebed62b1	Don't create partial matching words for synonyms, split words, phrases	2022-11-28 10:20:13 +01:00
Loïc Lecrenier	8284bd760f	Relax memory ordering of operations within the test CountingAlloc	2022-11-28 10:20:13 +01:00
Loïc Lecrenier	8d0ace2d64	Avoid creating a MatchingWord for words that exceed the length limit	2022-11-28 10:20:13 +01:00
Loïc Lecrenier	86c34a996b	Deduplicate matching words	2022-11-28 10:20:13 +01:00
Minh Pham	eba7af1d2c	Replace deprecated gh actions	2022-11-27 06:47:08 +07:00
Gregory Conrad	e0d24104a3	refactor: Rewrite another method chain to be more readable	2022-11-26 13:33:19 -05:00
Gregory Conrad	2db738dbac	refactor: rewrite method chain to be more readable	2022-11-26 13:26:39 -05:00
bors[bot]	84dd2e4df1	Merge #710 710: Update Clippy to use Rust Stable r=irevoire a=Kerollmops This PR changes the CI to use Rust stable for Clippy and Rustfmt. This way we will reduce the number of times we break the CI. [The version will only change every two months or so](https://www.whatrustisit.com/). Co-authored-by: Clément Renault <clement@meilisearch.com>	2022-11-24 15:57:04 +00:00
Clément Renault	3d06ea41ea	Keep a nightly for rustfmt Co-authored-by: Tamo <tamo@meilisearch.com>	2022-11-24 16:54:40 +01:00
Clément Renault	3958db4b17	Update the CI to use Rust Stable	2022-11-24 16:26:48 +01:00
Gregory Conrad	935a724c57	revert: Revert pass by reference API change	2022-11-24 10:08:23 -05:00
bors[bot]	914f8b118c	Merge #3119 3119: Dump tests r=Kerollmops a=irevoire Reenable the dump tests Co-authored-by: Tamo <tamo@meilisearch.com>	2022-11-24 10:17:38 +00:00
Gregory Conrad	ed29cceae9	perf: Prevent reindex in searchable set case when not needed	2022-11-23 22:33:06 -05:00
Gregory Conrad	bb9e33bf85	perf: Prevent reindex in searchable reset case when not needed	2022-11-23 22:01:46 -05:00
Gregory Conrad	7c0e544839	feat: Add all_obkv_to_json function	2022-11-23 21:18:58 -05:00
Colby Allen	6ecd486d1b	Bumps cargo_toml version to most up to date	2022-11-23 16:27:54 -07:00
Gregory Conrad	d19c8672bb	perf: limit reindex to when exact_attributes changes	2022-11-23 15:50:53 -05:00
Tamo	7b8641a7af	fix the dump tests The issue was linked to the fact that the debug implementation of the PhantomData wasn't the same between rust stable and rust nightly. This was causing an issue while snapshsotting the settings and this commit fix it by representing the settings as json which already ignores the PhantomData	2022-11-23 16:59:20 +01:00
bors[bot]	f509d81ec9	Merge #3100 3100: Add a dispatch to the publish binaries workflow r=Kerollmops a=curquiza Add `worklfow_dispatch` event to publish binaries workflow to allow the manually trigger Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-11-21 19:31:20 +00:00
Kerollmops	8d73ae80bb	Add a dispatch to the publish binaries workflow	2022-11-21 18:50:57 +01:00
bors[bot]	57c9f03e51	Merge #697 697: Fix bug in prefix DB indexing r=loiclec a=loiclec Where the batch's information was not properly updated in cases where only the proximity changed between two consecutive word pair proximities. Closes partially https://github.com/meilisearch/meilisearch/issues/3043 Co-authored-by: Loïc Lecrenier <loic.lecrenier@me.com>	2022-11-17 15:22:01 +00:00
bors[bot]	00129e112a	Merge #3041 3041: Add `workflow_dispatch` to flaky.yml r=irevoire a=curquiza To be able to run the job manual and don't wait for one week Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com>	2022-11-17 13:21:11 +00:00
bors[bot]	467e742bd1	Merge #702 702: Update version for the next release (v0.37.0) in Cargo.toml files r=ManyTheFish a=meili-bot ⚠️ This PR is automatically generated. Check the new version is the expected one before merging. Co-authored-by: curquiza <curquiza@users.noreply.github.com>	2022-11-17 12:54:27 +00:00
curquiza	cd5aaa3a9f	Update version for the next release (v0.37.0) in Cargo.toml files	2022-11-17 12:50:07 +00:00
bors[bot]	8ceb199dca	Merge #696 696: Fix Facet Indexing bugs r=Kerollmops a=loiclec 1. Handle keys with variable length correctly Closes (partially) https://github.com/meilisearch/meilisearch/issues/3042 This issue is now easily reproducible with the updated fuzz tests, which now generate keys with variable lengths. 2. Prevent adding facets to the database if their encoded value does not satisfy `valid_lmdb_key`. Closes (partially) https://github.com/meilisearch/meilisearch/issues/2743 This fixes an indexing failure when a document had a filterable attribute containing a value whose length is higher than ~500 bytes. For now, this fix is just meant to prevent crashes. Better handling of long values of filterable attributes will be handled in a separate PR. Co-authored-by: Loïc Lecrenier <loic.lecrenier@me.com>	2022-11-17 11:56:16 +00:00
Loïc Lecrenier	777eb3fa00	Add insta-snaps for test of bug 3043	2022-11-17 12:21:27 +01:00
Loïc Lecrenier	0caadedd3b	Make clippy happy	2022-11-17 12:17:53 +01:00
Loïc Lecrenier	ac3baafbe8	Truncate facet values that are too long before indexing them	2022-11-17 11:29:42 +01:00
Loïc Lecrenier	990a861241	Add test for indexing a document with a long facet value	2022-11-17 11:29:42 +01:00
Loïc Lecrenier	d95d02cb8a	Fix Facet Indexing bugs 1. Handle keys with variable length correctly This fixes https://github.com/meilisearch/meilisearch/issues/3042 and is easily reproducible with the updated fuzz tests, which now generate keys with variable lengths. 2. Prevent adding facets to the database if their encoded value does not satisfy `valid_lmdb_key`. This fixes an indexing failure when a document had a filterable attribute containing a value whose length is higher than ~500 bytes.	2022-11-17 11:29:42 +01:00
Loïc Lecrenier	f00108d2ec	Fix name of bug in reproduction test	2022-11-17 11:29:18 +01:00
Loïc Lecrenier	f7c8730d09	Fix bug in prefix DB indexing Where the batch's information was not properly updated in cases where only the proximity changed between two consecutive word pair proximities. Closes https://github.com/meilisearch/meilisearch/issues/3043	2022-11-17 11:29:18 +01:00
bors[bot]	46c275f0e4	Merge #3070 3070: Remove core and use engine r=Kerollmops a=curquiza Following the new team name Not mandatory since GitHub is doing redirection, but more consistent Co-authored-by: curquiza <clementine@meilisearch.com>	2022-11-16 16:40:44 +00:00
bors[bot]	a651397afc	Merge #685 685: ci: Use pre-compiled binaries for faster CI r=irevoire a=azzamsa # Pull Request ## Related issue Fixes #<issue_number> ## What does this PR do? - ... ## PR checklist Please check if your PR fulfills the following requirements: - [ ] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [ ] Have you read the contributing guidelines? - [ ] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: azzamsa <me@azzamsa.com>	2022-11-16 16:39:39 +00:00
curquiza	f4d4f313ea	Use the right new link	2022-11-16 17:21:35 +01:00
bors[bot]	2000db8453	Merge #701 701: Remove Hacktoberfest sections r=curquiza a=meili-bot _This PR is auto-generated._ Remove Hacktoberfest sections from CONTRIBUTING file. Co-authored-by: meili-bot <74670311+meili-bot@users.noreply.github.com>	2022-11-15 15:17:18 +00:00
meili-bot	92cc3550d8	Update CONTRIBUTING.md	2022-11-15 16:16:40 +01:00
bors[bot]	5e1fa53354	Merge #3055 3055: Remove Hacktoberfest sections r=curquiza a=meili-bot _This PR is auto-generated._ Remove Hacktoberfest sections from README and CONTRIBUTING files. Co-authored-by: meili-bot <74670311+meili-bot@users.noreply.github.com>	2022-11-15 14:29:27 +00:00
meili-bot	fe59a6f628	Update README.md	2022-11-15 15:27:18 +01:00
meili-bot	647b7a20e9	Update CONTRIBUTING.md	2022-11-15 15:27:17 +01:00
bors[bot]	cd3bca06e9	Merge #699 699: Force vendoring of LMDB even if a system version is available r=Kerollmops a=dureuill # Pull Request ## Related issue Related to https://github.com/meilisearch/meilisearch/issues/3017: will fix once ported to milli and meilisearch. ## What does this PR do? - Force using vendored version of LMDB - don't use lmdb master3 branch anymore: this is a bit of a side effect of using a tag instead of branch for heed as a dependency, but it is wanted anyway for now as lmdb master3 was more of an experiment - modifies CI to run `cargo check` on the release rather than the debug artifacts. This is an attempt to reduce the necessary disk space and avoid "out of space" failures. ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2022-11-15 11:09:20 +00:00
Louis Dureuil	87576cf26c	Perform cargo check on the release artifacts	2022-11-15 10:25:02 +01:00
funilrys	e81b349658	Fix linting issue.	2022-11-14 18:51:34 +01:00
Louis Dureuil	6dc6a5d874	Force using vendored version of LMDB - don't use lmdb master3 branch anymore	2022-11-14 17:17:51 +01:00
Clémentine Urquizar - curqui	a84fad5ce6	Add `workflow_dispatch` to flaky.yml	2022-11-14 10:10:00 +01:00
funilrys	0a102d601c	Update Task.created_at Indeed, before this patch we weren't considering the TaskContent::SetingsUpdate while trying to find the creation date.	2022-11-13 10:14:20 +01:00
funilrys	8a14f6f545	Add Task.processed_at.	2022-11-13 10:13:10 +01:00
funilrys	079357ee1f	Fix linting issues.	2022-11-12 20:57:27 +01:00
funilrys	06e7db7a1f	fixup! Extract the dates out of the dumpv4.	2022-11-12 18:28:23 +01:00
bors[bot]	9e189f5041	Merge #3015 3015: Replace deprecated set-output in GitHub actions r=curquiza a=funilrys # Pull Request This patch fixes #3011. This patch fixes the deprecation warning regarding the usage of `set-output`. This patch fixes the issues by switching the following format: ``` echo ::set-output name=[name]::[value] ``` into the following format: ``` echo "[name]=[value]" >> ${GITHUB_OUTPUT} ``` ## Related issue Fixes #3011 ## What does this PR do? - Fix CI/CD deprecation warnings. ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: funilrys <contact@funilrys.com> Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com>	2022-11-10 09:45:55 +00:00
Clémentine Urquizar - curqui	fe980f9e88	Update .github/workflows/publish-docker-images.yml	2022-11-10 10:40:21 +01:00
Clémentine Urquizar - curqui	32cd9e4852	Update .github/workflows/publish-docker-images.yml	2022-11-10 10:40:16 +01:00
Clémentine Urquizar - curqui	8d79f501f3	Update .github/workflows/publish-binaries.yml	2022-11-10 10:40:09 +01:00
Clémentine Urquizar - curqui	c05abc2b0d	Update .github/workflows/publish-binaries.yml	2022-11-10 10:40:04 +01:00
bors[bot]	e75829aded	Merge #694 694: Update version for the next release (v0.36.0) in Cargo.toml files r=curquiza a=meili-bot ⚠️ This PR is automatically generated. Check the new version is the expected one before merging. Co-authored-by: Kerollmops <Kerollmops@users.noreply.github.com>	2022-11-09 11:22:24 +00:00
Kerollmops	d00d2aab3f	Update version for the next release (v0.36.0) in Cargo.toml files	2022-11-09 11:03:09 +00:00
bors[bot]	f46a8ab2e2	Merge #693 693: use the lmdb-master.3 branch r=Kerollmops a=irevoire After investigating https://github.com/meilisearch/meilisearch/issues/3017, we found out that it was due to lmdb and that, without any code change on our side, bumping using the lmdb-master-3 branch fix our issues. But, we’re not really confident about what changed between the `mdb.master` and `mdb.master3` branches; thus this is a temporary change, and we hope we’ll be able to move to the new version of heed asap (either before the end of the pre-release or for the next release). -------- The bug is hard to reproduce; I can reproduce it 100% of the time on my archlinux personal computer. But on a scaleway archlinux bare-metal machine, it doesn’t reproduce. It’s flaky on our test suite, but `@loiclec` was able to write a minimal test that reproduces it every time on macOS. Basically, what happens is when there are multiple threads opening databases in a different directory at the same time. If there are 10 or more threads running at the same time, lmdb starts throwing the `Invalid argument (os error 22)` error for no reason, we believe. I would like to submit an issue to lmdb, but I don’t really have the time to write a test in C without heed currently. `@hyc,` if you want to take a look at it, here is the repo that reproduces the issue on macOS: https://github.com/irevoire/heed-bug Co-authored-by: Irevoire <tamo@meilisearch.com>	2022-11-09 09:42:38 +00:00
funilrys	a441fe5ae5	Remove unecessary line.	2022-11-08 21:18:24 +01:00
funilrys	7331da0410	Fix auto-formater issue. Indeed, my editor always fixes the format for me. That's why those 2 lines were changed.	2022-11-08 21:16:47 +01:00
funilrys	72c4db4553	Rewrite: ${GITHUB_OUTPUT} -> $GITHUB_OUTPUT.	2022-11-08 21:15:28 +01:00
bors[bot]	c3b75bbe5d	Merge #691 691: Update version for the next release (v0.35.1) in Cargo.toml files r=curquiza a=meili-bot ⚠️ This PR is automatically generated. Check the new version is the expected one before merging. Co-authored-by: Kerollmops <Kerollmops@users.noreply.github.com>	2022-11-08 15:31:50 +00:00
Irevoire	c7711daca3	use the lmdb-master.3 branch	2022-11-08 16:28:01 +01:00
bors[bot]	f18a4581f1	Merge #692 692: Update CONTRIBUTING.md r=Kerollmops a=curquiza Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com>	2022-11-08 14:51:30 +00:00
Clémentine Urquizar - curqui	8ce8bbcdfc	Update CONTRIBUTING.md	2022-11-08 15:49:45 +01:00
Kerollmops	bd12989610	Update version for the next release (v0.35.1) in Cargo.toml files	2022-11-08 14:31:39 +00:00
bors[bot]	24a298a83c	Merge #690 690: Fix soft deleted bug settings r=ManyTheFish a=Kerollmops Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-11-08 13:45:10 +00:00
bors[bot]	d85cd9bf1a	Merge #689 689: Handle non-finite floats consistently in filters r=irevoire a=dureuill # Pull Request ## Related issue Related meilisearch/meilisearch#3000 ## What does this PR do? ### User - Filters using `field = inf`, (or `infinite`, `NaN`) now match the value as a string rather than returning an internal error. - Filters using `field < inf` (or other comparison operators) now return an invalid_filter error rather than returning an internal error, much like when using `field < aaa`. ### Implementation - Add new `NonFiniteFloat` error variants to the filter-parser errors - Add `Token::parse_as_finite_float` that can fail both when the string is not a float and when the float is not finite - Refactor `Filter::inner_evaluate` to always use `parse_as_finite_float` instead of just `parse` - Add corresponding tests ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Louis Dureuil <louis@meilisearch.com>	2022-11-08 13:24:38 +00:00
Kerollmops	37b3c5c323	Fix transform to use all_documents and ignore soft_deleted documents	2022-11-08 14:23:16 +01:00
Kerollmops	1b1ad1923b	Add a test to check that we take care of soft deleted documents	2022-11-08 14:23:14 +01:00
Louis Dureuil	a836b8e703	tests: Tests filter with non-finite floats	2022-11-08 13:56:55 +01:00
Louis Dureuil	3328560788	fix: allow filters on = inf, = NaN, return InvalidFilter for < inf, < NaN Fixes meilisearch/meilisearch#3000	2022-11-08 13:27:15 +01:00
bors[bot]	cf76ec7b37	Merge #673 673: Add clippy job r=ManyTheFish a=unvalley # Pull Request ## Related issue Fixes #231 ## What does this PR do? - fix some clippy errors remain - add clippy job to CI (I set `nightly` as toolchain) ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Co-authored-by: unvalley <kirohi.code@gmail.com>	2022-11-08 09:43:26 +00:00
unvalley	abf1cf9cd5	Fix clippy errors	2022-11-04 09:27:46 +09:00
unvalley	b09676779d	Use nightly for clippy and remove conflict mistake	2022-11-04 09:13:01 +09:00
unvalley	70465aa5ce	Execute cargo fmt	2022-11-04 08:59:58 +09:00
unvalley	3009981d31	Fix clippy errors Add clippy job Add clippy job to CI	2022-11-04 08:58:14 +09:00
unvalley	401e956128	Add clippy job Add clippy job to CI	2022-11-04 08:58:12 +09:00
azzamsa	48eafc546f	ci: Use pre-compiled binaries for faster CI	2022-11-04 00:03:53 +07:00
bors[bot]	6add470805	Merge #659 659: Fix clippy error to add clippy job on Ci r=Kerollmops a=unvalley ## Related PR This PR is for #673 ## What does this PR do? - ~~add `Run Clippy` job to CI (rust.yml)~~ - apply `cargo clippy --fix` command - fix some `cargo clippy` error manually (but warnings still remain on tests) ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Co-authored-by: unvalley <kirohi.code@gmail.com> Co-authored-by: unvalley <38400669+unvalley@users.noreply.github.com>	2022-11-03 15:24:38 +00:00
unvalley	13175f2339	refactor: match for filterCondition	2022-11-03 17:34:33 +09:00
funilrys	953b2ec438	fixup! Extract the dates out of the dumpv4.	2022-11-02 17:49:37 +01:00
bors[bot]	1a1ad8a792	Merge #679 679: Bump Swatinem/rust-cache from 2.0.0 to 2.0.1 r=curquiza a=dependabot[bot] Bumps [Swatinem/rust-cache](https://github.com/Swatinem/rust-cache) from 2.0.0 to 2.0.1. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/Swatinem/rust-cache/releases">Swatinem/rust-cache's releases</a>.</em></p> <blockquote> <h2>v2.0.1</h2> <ul> <li>Primarily just updating dependencies to fix GitHub deprecation notices.</li> </ul> </blockquote> </details> <details> <summary>Changelog</summary> <p><em>Sourced from <a href="https://github.com/Swatinem/rust-cache/blob/master/CHANGELOG.md">Swatinem/rust-cache's changelog</a>.</em></p> <blockquote> <h2>2.0.1</h2> <ul> <li>Primarily just updating dependencies to fix GitHub deprecation notices.</li> </ul> </blockquote> </details> <details> <summary>Commits</summary> <ul> <li><a href="`22c9328bcb`"><code>22c9328</code></a> 2.0.1</li> <li><a href="`d4d463bd9b`"><code>d4d463b</code></a> bump deps and rebuild</li> <li><a href="`c4652c677c`"><code>c4652c6</code></a> Update <code>`@actions/core</code>` (<a href="https://github-redirect.dependabot.com/Swatinem/rust-cache/issues/83">#83</a>)</li> <li><a href="`76686c56f2`"><code>76686c5</code></a> docs: Fix github workflows directory (<a href="https://github-redirect.dependabot.com/Swatinem/rust-cache/issues/79">#79</a>)</li> <li><a href="`1b43d2f2c3`"><code>1b43d2f</code></a> remove outdated versioning note</li> <li><a href="`20b9201e8a`"><code>20b9201</code></a> bump cargo hash</li> <li><a href="`0d72e5f9a0`"><code>0d72e5f</code></a> revert explicit dir close</li> <li><a href="`86531941c2`"><code>8653194</code></a> Merge branch 'master' of <a href="https://github.com/Swatinem/rust-cache">https://github.com/Swatinem/rust-cache</a></li> <li><a href="`be4be3720d`"><code>be4be37</code></a> explicitly close dir handles, add more logging, cleanups</li> <li><a href="`213334cd98`"><code>213334c</code></a> cargo update</li> <li>Additional commits viewable in <a href="https://github.com/Swatinem/rust-cache/compare/v2.0.0...v2.0.1">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=Swatinem/rust-cache&package-manager=github_actions&previous-version=2.0.0&new-version=2.0.1)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) You can trigger a rebase of this PR by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) </details> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2022-11-02 12:38:05 +00:00
dependabot[bot]	4492605a78	Bump Swatinem/rust-cache from 2.0.0 to 2.0.1 Bumps [Swatinem/rust-cache](https://github.com/Swatinem/rust-cache) from 2.0.0 to 2.0.1. - [Release notes](https://github.com/Swatinem/rust-cache/releases) - [Changelog](https://github.com/Swatinem/rust-cache/blob/master/CHANGELOG.md) - [Commits](https://github.com/Swatinem/rust-cache/compare/v2.0.0...v2.0.1) --- updated-dependencies: - dependency-name: Swatinem/rust-cache dependency-type: direct:production update-type: version-update:semver-patch ... Signed-off-by: dependabot[bot] <support@github.com>	2022-11-01 10:19:45 +00:00
funilrys	09e71fdeb6	Replace deprecated set-output in GitHub actions This patch fixes #3011. This patch fixes the depracation warning regarding the usage of `set-output`. This patch fixes the issues by switching the following format: ``` echo ::set-output name=[name]::[value] ``` into the following format: ``` echo "[name]=[value]" >> ${GITHUB_OUTPUT} ```	2022-10-31 22:28:01 +01:00
bors[bot]	fe5a0219e1	Merge #677 677: run the tests in all workspaces r=curquiza a=irevoire With #676 I noticed the tests were not running in any of our sub crates. Most of our sub crates didn't includes any tests though. But the filter-parser did and we're lucky we never broke these one without noticing 😁 Co-authored-by: Irevoire <tamo@meilisearch.com>	2022-10-31 18:05:04 +00:00
funilrys	ab3056cc66	Extract the dates out of the dumpv4. This patch possibly fixes #2987. This patch introduces a way to fill the IndexMetadata.created_at and IndexMetadata.updated_at keys from the tasks events. This is done by reading the creation date of the first event (created_at) and the creation date of the last event (updated_at).	2022-10-31 18:58:14 +01:00
Irevoire	5ff066c3e7	run the tests in all workspaces	2022-10-31 18:38:48 +01:00
bors[bot]	6770eb2a87	Merge #676 676: chore: added `IN`,`NOT IN` to `invalid_filter` msg r=irevoire a=Pranav-yadav # Pull Request ## Related issue `Fixes` https://github.com/meilisearch/meilisearch/issues/3004 ## What does this PR do? - Improves correct error msg in response ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Pranav Yadav <Pranavyadav3912@gmail.com>	2022-10-31 17:29:24 +00:00
unvalley	0d43ddbd85	Update filter-parser/src/lib.rs Co-authored-by: Tamo <irevoire@protonmail.ch>	2022-11-01 01:32:54 +09:00
bors[bot]	c7caadb54e	Merge #3001 3001: Implement Uuid codec for heed r=Kerollmops a=elbertronnie # Pull Request ## Related issue Fixes #2984 ## What does this PR do? - Created a new heed codec for uuid::Uuid named as UuidCodec - Replaced SerdeBincode\<Uuid\> with UuidCodec - Removed the TODO in code associated with this issue ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Elbert Ronnie <elbert.ronniep@gmail.com>	2022-10-31 16:13:20 +00:00
Pranav Yadav	3950ec8d3c	chore: update tests for `invalid_filter` msg	2022-10-31 15:41:49 +00:00
Pranav Yadav	3b35ebda50	chore: added `IN`,`NOT IN` to `invalid_filter` msg	2022-10-31 15:01:14 +00:00
bors[bot]	2254bbf3bd	Merge #3002 3002: Fix dump import without instance uid r=Kerollmops a=irevoire When creating a dump without any instance-uid (that can happen if you’ve always run meilisearch with the `--no-analytics` flag), you could get an error when trying to load the dump. Co-authored-by: Irevoire <tamo@meilisearch.com>	2022-10-31 12:58:37 +00:00
Elbert Ronnie	0219ef25fe	Moved the struct UuidCodec to a new file	2022-10-31 12:25:19 +05:30
Irevoire	510afda590	remove unused import	2022-10-30 20:05:20 +01:00
Irevoire	fea9fdcd7e	fix the dump reader process when no instance-uid was specified	2022-10-30 20:00:27 +01:00
bors[bot]	4bcfd14a45	Merge #675 675: Deleted empty files r=Kerollmops a=SKVKPandey # Pull Request ## Related issue Fixes #674 ## What does this PR do? Delete empty files: - `milli/src/heed_codec/facet/facet_string_level_zero_value_codec.rs` - `milli/src/heed_codec/facet/facet_string_zero_bounds_value_codec.rs` ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Shashank Kashyap <50551759+SKVKPandey@users.noreply.github.com>	2022-10-30 07:09:30 +00:00
Shashank Kashyap	a07f0a4a43	Delete facet_string_zero_bounds_value_codec.rs	2022-10-30 08:59:04 +05:30
Shashank Kashyap	2dec6e86e9	Delete facet_string_level_zero_value_codec.rs	2022-10-30 08:58:36 +05:30
Elbert Ronnie	3911fd64b5	Implement Uuid codec for heed	2022-10-30 03:27:30 +05:30
bors[bot]	c965200010	Merge #664 664: Fix phrase search containing stop words r=ManyTheFish a=Samyak2 # Pull Request This a WIP draft PR I wanted to create to let other potential contributors know that I'm working on this issue. I'll be completing this in a few hours from opening this. ## Related issue Fixes #661 and towards fixing meilisearch/meilisearch#2905 ## What does this PR do? - [x] Change Phrase Operation to use a `Vec<Option<String>>` instead of `Vec<String>` where `None` corresponds to a stop word - [x] Update all other uses of phrase operation - [x] Update `resolve_phrase` - [x] Update `create_primitive_query`? - [x] Add test ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Co-authored-by: Samyak S Sarnayak <samyak201@gmail.com> Co-authored-by: Samyak Sarnayak <samyak201@gmail.com>	2022-10-29 13:42:52 +00:00
unvalley	d55f0e2e53	Execute cargo fmt	2022-10-28 23:42:23 +09:00
unvalley	d53a80b408	Fix clippy error	2022-10-28 23:41:35 +09:00
Samyak Sarnayak	ecb88143f9	Run cargo fmt	2022-10-28 19:37:02 +05:30
Samyak Sarnayak	03eb5d87c1	Only call plane_sweep on subgroups when 2 or more are present	2022-10-28 19:32:05 +05:30
unvalley	a1d7ed1258	fix clippy error and remove clippy job from ci Remove clippy job Fix clippy error type_complexity Restore ambiguous change	2022-10-28 22:33:50 +09:00
unvalley	f3c0b05ae8	Fix rust fmt	2022-10-28 09:32:31 +09:00
bors[bot]	dd1011ba76	Merge #2995 2995: merge the settings and do one indexation at the end r=irevoire a=irevoire Co-authored-by: Irevoire <tamo@meilisearch.com>	2022-10-27 21:24:21 +00:00
bors[bot]	20258461a8	Merge #2981 #2996 2981: Move index swap error handling from meilisearch-http to index-scheduler r=irevoire a=loiclec And make index_not_found error asynchronous, since we can't know whether the index will exist by the time the index swap task is processed. Improve the index-swap test to verify that future tasks are not swapped and to test the new error messages that were introduced. ## Related issue https://github.com/meilisearch/meilisearch/issues/2973 2996: Get rids of the unecessary tasks when an index_uid is specified r=Kerollmops a=irevoire Co-authored-by: Loïc Lecrenier <loic.lecrenier@me.com> Co-authored-by: Tamo <tamo@meilisearch.com> Co-authored-by: Irevoire <tamo@meilisearch.com>	2022-10-27 19:11:23 +00:00
Tamo	87cac158c4	Update index-scheduler/src/batch.rs	2022-10-27 18:08:21 +02:00
Tamo	c9f89d38e3	Merge branch 'main' into index-swap-error-handling	2022-10-27 18:06:45 +02:00
Irevoire	01687c87a2	Get rids of the unecessary tasks when an index_uid is specified	2022-10-27 18:00:04 +02:00
unvalley	f4ec1abb9b	Fix all clippy error after conflicts	2022-10-27 23:58:13 +09:00
Irevoire	313f204f39	merge the settings and do one indexation at the end	2022-10-27 16:38:21 +02:00
bors[bot]	d16ea755d8	Merge #2982 2982: Adapt task queries to account for special index swap rules r=irevoire a=loiclec # Pull Request ## Related issue Fixes https://github.com/meilisearch/meilisearch/issues/2970 ## What does this PR do? - Replace the `get_tasks` method with a `get_tasks_from_authorized_indexes` which returns the list of tasks matched by the query from the point of view of the user. That is, it takes into consideration the list of authorised indexes as well as the special case of `IndexSwap` which should not be returned if an index_uid is specified or if any of its associated indexes are not authorised. - Adapt the code in other places following this change - Add some tests - Also the method `get_task_ids_from_authorized_indexes` now takes a read transaction as argument. This is because we want to make sure that the implementation of `get_tasks_from_authorized_indexes` only uses one read transaction. Otherwise, we could (1) get a list of task ids matching the query, then (2) one of these task ids is deleted by a taskDeletion task, and finally (3) we try to get the `Task`s associated with each returned task ids, and get a `CorruptedTaskQueue` error. Co-authored-by: Loïc Lecrenier <loic.lecrenier@me.com>	2022-10-27 14:28:04 +00:00
Loïc Lecrenier	8152ab5dfc	Revert change in initialisation of TempDir for index scheduler tests	2022-10-27 16:26:17 +02:00
Loïc Lecrenier	8dd7942656	Cargo fmt	2022-10-27 16:24:09 +02:00
Loïc Lecrenier	2c31d7c50a	Apply review suggestions	2022-10-27 16:24:08 +02:00
bors[bot]	b76f0ace26	Merge #2993 2993: Reconsider the Windows tests r=irevoire a=Kerollmops This PR removes the `ignore` cfg on top of a lot of our tests. Now that we reworked the index scheduler we can make them pass again! Fixes #2038, fixes #1966. Co-authored-by: Clément Renault <clement@meilisearch.com>	2022-10-27 13:41:04 +00:00
bors[bot]	5b535a82ea	Merge #2991 2991: Update version for the next release (v0.30.0) in Cargo.toml files r=Kerollmops a=meili-bot ⚠️ This PR is automatically generated. Check the new version is the expected one before merging. Co-authored-by: curquiza <curquiza@users.noreply.github.com>	2022-10-27 13:12:31 +00:00
Clément Renault	e67673bd12	Ingore the dumps v1 test on Windows	2022-10-27 14:34:45 +02:00
Clément Renault	44d6f3e7a0	Reconsider the Windows tests	2022-10-27 13:50:05 +02:00
curquiza	68f80dbacf	Update version for the next release (v0.30.0) in Cargo.toml files	2022-10-27 11:35:44 +00:00
bors[bot]	08db774699	Merge #2990 2990: isolate the search in another task r=Kerollmops a=irevoire In case there is a failure on milli's side that should avoid blocking the tokio main thread Co-authored-by: Irevoire <tamo@meilisearch.com>	2022-10-27 11:29:22 +00:00
Irevoire	4d9e9f4a9d	isolate the search in another task In case there is a failure on milli's side that should avoid blocking the tokio main thread	2022-10-27 13:12:42 +02:00
Loïc Lecrenier	4f4fc20acf	Make clippy happy	2022-10-27 13:00:30 +02:00
Loïc Lecrenier	78ffa00f98	Move index swap error handling from meilisearch-http to index-scheduler And make index_not_found error asynchronous, since we can't know whether the index will exist by the time the index swap task is processed. Improve the index-swap test to verify that future tasks are not swapped and to test the new error messages that were introduced.	2022-10-27 11:45:38 +02:00
Loïc Lecrenier	7b93ba40bd	Reimplement task queries to account for special index swap rules	2022-10-27 11:44:51 +02:00
bors[bot]	b44cc62320	Merge #2763 2763: Index scheduler r=Kerollmops a=irevoire Fix https://github.com/meilisearch/meilisearch/issues/2725 - [x] Durability of the tasks once an answer has been sent to the user. - [x] Fix the analytics - [x] Disable the auto-batching system. - [x] Make sure the task scheduler run if there are tasks to process. - [x] Auto-batching of enqueued tasks: - [x] Do not batch operations from two different indexes. - [x] Document addition. - [x] Document updates. - [x] Settings. - [x] Document deletion. - [x] Make sure that we only merge batches with the same index-creation rights: - [x] the batch either starts with a `yes` - [x] [we only batch `no`s together and stop batching when we encounter a `yes`](https://www.youtube.com/watch?v=O27mdRvR1GY) - [x] Unify the logic about `false` and `true` index creation rights. - [ ] Execute all batch kind: - [x] Import dumps at startup time. - [x] Export dumps i.e. export the tasks queue. - [x] Document addition - [x] Document update - [x] Document deletion. - [x] Clear all documents. - [x] Update the settings of an index. - [ ] Merge multiple settings into a single one. - [x] Index update e.g. Create an Index, change an index primary key, delete an index. - [x] Cancel enqueued or processing tasks (with filters) (don't count tasks from forbidden indexes) (can't cancel a task with a higher or equal task_id than your own). - [x] Delete processed tasks from the task store (with filters) (don't count tasks from forbidden indexes) (can't flush a task with a higher or equal task_id than your own) - [x] Document addition + settings - [x] Document addition + settings + clear all documents - [x] anything + index deletion - [x] Snapshot - [x] Make the `SnapshotCreation` task visible. - [x] Snapshot tasks are scheduled by a detached thread. - [x] Only include update files that are useful. - [x] Check that statuses and details are correctly set. (ie; if you enqueue a `documentAddition`, is the `documentReceived` well set?) - [x] Prioritize and reorder tasks i.e. Index deletion, Delete all the documents. - [x] Always accept new tasks without blocking. - [x] Fairly share the loads over the different indexes e.g. Always process the index queue with the lowest id. - [x] Easily testable. - [x] Well tested i.e. tasks reordering, tasks prioritizing, use atomic barriers to block the tasks for tests. - [x] Dump - [x] Serialize the uuid as string in the keys - [x] Create a dump crate with getters and setters - [x] Serialize the API key in the dump task - [x] Get the instance-uuid in the dump task - [x] List and filter tasks: - [x] Paginate the tasks. - [x] Filter by index name. - [x] Filter on the status, the enqueued, processing, and finished tasks. - [x] Filter on the type of task. - [x] Check that it works in `meilisearch-http`. - [x] Think about [the index wrapper](`2c4c14caa8/index/src/updates.rs (L269)`) and probably move or remove it. - [x] Reduce the amount of copy/paste for the batched operations by creating a sub-enum for the `Batch` enum. - [x] Move the `IndexScheduler` in the lib.rs file. - [x] Think about the `MilliError` type and probably remove it. - [x] Remove the `index` crate entirely - [x] Remove the `Kind` type from the `TaskView` and introduce another type, remove the `<Kind as FromStr>`. - [x] Once the point above is done; remove the unreachable variant from the autobatchingkind - [x] Rename the `Settings` task `Kind` to `SettingsUpdate` - [x] Rename the `DumpExport` task `Kind` to `DumpExport` - [x] Path the error message when deserializing a `Kind` and `Status`. - [x] Check the version file when starting. - [x] Copy the version file when creating snapshots. --------- Once everything above is done; - [ ] Check what happens with the update files i.e. when are they deleted. - [ ] When a TaskDeletion occurs - [ ] When a TaskCancelation - [ ] When a task is finished - [ ] When a task fails - [ ] When importing a dump forward the date to milli - [ ] Add tests for the snapshots. - [ ] Look at all the places where we put _TODOs_. - [ ] Rename a bunch of things, see https://github.com/meilisearch/meilisearch/pull/2917 - [ ] Ensure that when compiling meilisearch-http with `no-default-features` it doesn’t pull lindera etc - [ ] Run a bunch of operations in a `tokio::spawn_blocking` - [ ] The search requests - [ ] Issue to create once this is merged: - [ ] Realtime progressing status e.g. Websocket events (optional). - [ ] Implement an `Uuid` codec instead of using a `Bincode<Uuid>`. - [ ] Handle the dump-v1 - [ ] When importing a dump v1 we could iterate over the whole task queue to find the creation and last update date - [ ] When importing a dump v2 we could iterate over the whole task queue to find the creation and last update date - [ ] When importing a dump v3 we could iterate over the whole task queue to find the creation and last update date - [ ] When importing a dump v4 we could iterate over the whole task queue to find the creation and last update date - [ ] When importing a dump v5 we could iterate over the whole task queue to find the creation and last update date Co-authored-by: Tamo <tamo@meilisearch.com> Co-authored-by: Kerollmops <clement@meilisearch.com> Co-authored-by: Loïc Lecrenier <loic@meilisearch.com>	2022-10-27 09:38:00 +00:00
Clément Renault	fae17ed590	Enable the authentication tests on Windows again	2022-10-27 11:35:24 +02:00
Clément Renault	7e355958e0	Await the last insert task	2022-10-27 11:35:24 +02:00
Irevoire	8bc602a7dd	makes clippy happy	2022-10-27 11:35:23 +02:00
Irevoire	6c2ecec4d0	fix the return of the task cancelation and task deletion	2022-10-27 11:35:23 +02:00
Irevoire	6280bd51a9	actually fix the test and the swap_indexes name resolution	2022-10-27 11:35:23 +02:00
Irevoire	54d0aff4cf	ignore a strange test that works on my machine but not on the ci	2022-10-27 11:35:23 +02:00
Irevoire	b804cba4ca	try to debug the ci	2022-10-27 11:35:23 +02:00
Irevoire	3cf8aaa4d0	reformat	2022-10-27 11:35:23 +02:00
Irevoire	225405bb0d	ignore the dump tests	2022-10-27 11:35:22 +02:00
Kerollmops	0dd8e00929	Reapply #2601	2022-10-27 11:35:22 +02:00
Irevoire	a99ddf85f7	fix clippy once again	2022-10-27 11:35:22 +02:00
Irevoire	7307c4dacd	fix clippy	2022-10-27 11:35:22 +02:00
Irevoire	6aa816d96a	use meili-snap in the dump	2022-10-27 11:35:22 +02:00
Irevoire	866a3676eb	reupload the test fix for the dump	2022-10-27 11:35:22 +02:00
Kerollmops	fa84eae0f1	Insta review and fix insta snapshots	2022-10-27 11:35:21 +02:00
Irevoire	33996071ea	fix clippy from the CI	2022-10-27 11:35:21 +02:00
Irevoire	953055e3d7	bump milli	2022-10-27 11:35:21 +02:00
Kerollmops	7c908fadcf	Remove a useless clippy silence	2022-10-27 11:35:21 +02:00
Irevoire	07d39776f9	fix clippy _once again_	2022-10-27 11:35:21 +02:00
Irevoire	3979c9f02b	fix all the dump snasphots	2022-10-27 11:35:20 +02:00
Irevoire	8ec3681cf8	fix clippy part1	2022-10-27 11:35:20 +02:00
Kerollmops	2ba5e3b519	Clean up some code	2022-10-27 11:35:20 +02:00
Kerollmops	861a07792e	Remove useless task module	2022-10-27 11:35:20 +02:00
Kerollmops	ee6597da60	Fix all the tests	2022-10-27 11:35:20 +02:00
Kerollmops	314b89ca30	Fix insta snapshots	2022-10-27 11:35:20 +02:00
Irevoire	8ebb49d1b1	bump milli	2022-10-27 11:35:19 +02:00
Irevoire	0ba6253eed	fix the sort	2022-10-27 11:35:19 +02:00
Irevoire	e8cd571820	try to convert the OsStr to a rust string to fix the sort	2022-10-27 11:35:19 +02:00
Clément Renault	4f955e68b3	Apply suggestions from code review	2022-10-27 11:35:19 +02:00
Irevoire	6c98752922	move the commit before the insertion in the map	2022-10-27 11:35:19 +02:00
Irevoire	4e1b6b514e	update reviewer change	2022-10-27 11:35:19 +02:00
Irevoire	64e55b4db9	fix the index creation. When an index is being created we insert it in the index_map straight away to avoid someone else from trying to re-open it. The definitive fix should be made on milli's side	2022-10-27 11:35:18 +02:00
Loïc Lecrenier	9b43528bbb	Update test after fixing bug in index swap	2022-10-27 11:35:18 +02:00
Loïc Lecrenier	e641d08846	Cargo fmt	2022-10-27 11:35:18 +02:00
Loïc Lecrenier	36c9f05998	Revert "Display more than one indexUid in a task view if necessary" This reverts commit `1f2e253bb6`.	2022-10-27 11:35:18 +02:00
Loïc Lecrenier	3b158bb966	Return invalid API key error in /swap-indexes	2022-10-27 11:35:18 +02:00
Loïc Lecrenier	08b5123380	Display more than one indexUid in a task view if necessary	2022-10-27 11:35:17 +02:00
Loïc Lecrenier	1f75caae88	Fix a few index swap bugs. 1. Details of the indexSwap task 2. Query tasks with type=indexUid 3. Synchronous error message for multiple index not found	2022-10-27 11:35:17 +02:00
Irevoire	a16604af80	fix all the tests	2022-10-27 11:35:17 +02:00
Irevoire	1d014a538e	comment out a test that makes the CI crash	2022-10-27 11:35:17 +02:00
Irevoire	29bdcb880c	update the snapshot	2022-10-27 11:35:17 +02:00
Irevoire	a3fc0d3bd9	Fix the last regression	2022-10-27 11:35:17 +02:00
Kerollmops	2de8a0711a	Cargo insta test/review	2022-10-27 11:35:16 +02:00
Kerollmops	2f577b6fcd	Patch the IndexScheduler in meilisearch-http to use the options struct	2022-10-27 11:35:16 +02:00
Tamo	eccbdb74cf	remove useless print Co-authored-by: Clément Renault <clement@meilisearch.com>	2022-10-27 11:35:16 +02:00
Irevoire	033794d209	add tests for the task deletion and task cancelation	2022-10-27 11:35:16 +02:00
Irevoire	a85d5b4981	test the details of all tasks type	2022-10-27 11:35:16 +02:00
Kerollmops	71b50853dc	Introduce an options struct to create the IndexScheduler	2022-10-27 11:35:16 +02:00
Kerollmops	7074872a78	cargo insta accept	2022-10-27 11:35:15 +02:00
Kerollmops	035e8eeff5	Clean-up some TODOs	2022-10-27 11:35:15 +02:00
Kerollmops	e35fe33712	Fix some bugs with files	2022-10-27 11:35:15 +02:00
Kerollmops	4736e00253	Handle the CLI options related to snapshots	2022-10-27 11:35:15 +02:00
Kerollmops	942b7c338b	Compress the snapshot in a tarball	2022-10-27 11:35:15 +02:00
Kerollmops	4cafc63561	Reintroduce the versioning functions	2022-10-27 11:35:14 +02:00
Kerollmops	89e127e4f4	Declare the auth path in the index scheduler	2022-10-27 11:35:14 +02:00
Kerollmops	eec43ec953	Implement a first version of the snapshots	2022-10-27 11:35:14 +02:00
Kerollmops	c063f154fb	Add the snapshots directory path to the IndexScheduler	2022-10-27 11:35:14 +02:00
Kerollmops	e0548e42e7	Rename the Snapshot task into SnapshotCreation	2022-10-27 11:35:14 +02:00
Kerollmops	4d43a9f5b1	Rename the index-scheduler module into insta_snapshot	2022-10-27 11:35:14 +02:00
Kerollmops	901c405919	Fix the inta-snapshot typos in the tests	2022-10-27 11:35:13 +02:00
Kerollmops	c641888a23	Patch the delete and cancel tasks routes	2022-10-27 11:35:13 +02:00
Loïc Lecrenier	6db90ba6cc	Make sure that we don't delete or cancel future tasks This should already have been the case before, but there is no harm in adding another check.	2022-10-27 11:35:13 +02:00
Irevoire	e0821ad4b0	remove an useless dbg	2022-10-27 11:35:13 +02:00
Irevoire	61f0940f8c	fix an issue with the dates	2022-10-27 11:35:13 +02:00
Irevoire	241300d2d8	add more naive tests around the document addition + remove the old unused snapshot files	2022-10-27 11:35:13 +02:00
Irevoire	570b2d1167	add some naive document addition tests	2022-10-27 11:35:12 +02:00
Loïc Lecrenier	d92425658e	Add index scheduler tests for task cancelation	2022-10-27 11:35:12 +02:00
Irevoire	12669bf07c	rename received_documents_ids to matched_documents	2022-10-27 11:35:12 +02:00
Loïc Lecrenier	16fac10074	Fix crash when batching an index swap task containing 0 swaps	2022-10-27 11:35:12 +02:00
Irevoire	0aca5e84b9	rename received_document_ids to matched_documents in the DocumentDeletion task type (reimplementation of #2826 )	2022-10-27 11:35:12 +02:00
Irevoire	7ed3f00b1e	reformat	2022-10-27 11:35:12 +02:00
Irevoire	9c00b159ba	fix clippy	2022-10-27 11:35:11 +02:00
Irevoire	7e52f1effb	remove a lot of unecessary clone and ref	2022-10-27 11:35:11 +02:00
Loïc Lecrenier	4d25c159e6	Apply code review suggestions	2022-10-27 11:35:11 +02:00
Loïc Lecrenier	e9cd6cbbee	Revert implementation of `get_status` to query only the database	2022-10-27 11:35:11 +02:00
Loïc Lecrenier	424202d773	Pause the index scheduler for one second when a fatal error occurs	2022-10-27 11:35:11 +02:00
Loïc Lecrenier	4a35eb9849	Fix (hopefully) queries that include processing tasks	2022-10-27 11:35:11 +02:00
Loïc Lecrenier	493a8cff31	Adjust task details correctly following index swap	2022-10-27 11:35:10 +02:00
Loïc Lecrenier	4de445d386	Start testing unexpected errors and panics in index scheduler	2022-10-27 11:35:10 +02:00
Loïc Lecrenier	e3848b5f28	Add assert method to verify validity of index scheduler state	2022-10-27 11:35:10 +02:00
Irevoire	ecf4e43b3d	rename the dumpExport to dumpCreation	2022-10-27 11:35:10 +02:00
Irevoire	3ea489421e	move the error types to meilisearch-http	2022-10-27 11:35:10 +02:00
Loïc Lecrenier	2808be9d45	Fix the /swap-indexes route API 1. payload 2. error messages 3. auth errors	2022-10-27 11:35:10 +02:00
Loïc Lecrenier	92c41f0ef6	meili-snap: get the test name from the name of the function	2022-10-27 11:35:09 +02:00
Loïc Lecrenier	1214a68a41	Only store full snapshots if env variable is set to true	2022-10-27 11:35:09 +02:00
Irevoire	8a23e707c1	fix the task view and forward the task db size	2022-10-27 11:35:09 +02:00
Irevoire	eb4bdde432	fix clippy	2022-10-27 11:35:09 +02:00
Irevoire	735a5da257	reformat	2022-10-27 11:35:09 +02:00
Irevoire	1d04ce611d	remove ununsed function	2022-10-27 11:35:08 +02:00
Irevoire	e9055f5572	fix clippy	2022-10-27 11:35:08 +02:00
Irevoire	874499a2d2	fix all the snapshots	2022-10-27 11:35:08 +02:00
Irevoire	ecdcbf350f	update all the snapshots with the new kind name	2022-10-27 11:35:08 +02:00
Irevoire	de7d4200d8	fix the snapshot tests of the dump after renaming a bunch of kinds	2022-10-27 11:35:08 +02:00
Irevoire	2c0fde4a1a	ignore the snapshot test	2022-10-27 11:35:08 +02:00
Irevoire	e2ce8f2d32	remove the public DocumentClear variant	2022-10-27 11:35:07 +02:00
Irevoire	c8ee453b6c	fix the autobatched document deletion	2022-10-27 11:35:07 +02:00
Irevoire	f6963f9662	ensure the indexUid is valid in most cases	2022-10-27 11:35:07 +02:00
Irevoire	a8de5368e5	fix the index creation in case an index already exists	2022-10-27 11:35:07 +02:00
Irevoire	b3265a8e1f	ensure the index_uid is valid when creating an index	2022-10-27 11:35:07 +02:00
Irevoire	9bb2e3c790	fix the failed document addition with a primary key	2022-10-27 11:35:07 +02:00
Irevoire	cb48a02f94	fix the invalid index uid errors	2022-10-27 11:35:06 +02:00
Irevoire	99144b1419	fix most content file error	2022-10-27 11:35:06 +02:00
Irevoire	e107f1b282	fix the payload too large error	2022-10-27 11:35:06 +02:00
Irevoire	1bef5d119d	fix the api keys for the tasks route	2022-10-27 11:35:06 +02:00
Irevoire	ca4234b445	fix the deletion of the data.ms in case of failure	2022-10-27 11:35:06 +02:00
Irevoire	8d1408c65e	fix the import of the dumpv4&v5 when there is no instance-uid + rename the Kind+KindWithContent+Details variant for the DocumentImport and the Setting	2022-10-27 11:35:05 +02:00
Irevoire	131fe30934	fix the error messages and the index stats	2022-10-27 11:35:05 +02:00
Irevoire	50386921df	fix the index creation	2022-10-27 11:35:05 +02:00
Clément Renault	32cfac0cfd	Sort the TOML dependencies	2022-10-27 11:35:05 +02:00
Clément Renault	80b2e70ee7	Introduce a rustfmt file	2022-10-27 11:35:05 +02:00
Clément Renault	52e858a588	Reapply #2890	2022-10-27 11:34:18 +02:00
Clément Renault	8b0427f0c4	Reapply #2839	2022-10-27 11:34:18 +02:00
Clément Renault	fce0996e17	Reapply #2819	2022-10-27 11:34:18 +02:00
Clément Renault	2a7ef3b352	Reapply #2830	2022-10-27 11:34:18 +02:00
Clément Renault	ce4dcf47f0	Reapply #2773	2022-10-27 11:34:18 +02:00
Clément Renault	ca8c922f35	Reapply #2727	2022-10-27 11:34:18 +02:00
Clément Renault	4c42130ec7	Remove once for all the meilisearch-lib crate	2022-10-27 11:34:17 +02:00
Clément Renault	788262e588	Fix final compilation	2022-10-27 11:34:17 +02:00
Clément Renault	61edcd585a	Fix the new config file with the index scheduler	2022-10-27 11:34:17 +02:00
Clément Renault	72ec4ce96b	Fix allow_index_creation useless field	2022-10-27 11:34:17 +02:00
Clément Renault	75857bf476	Fix the insta tests	2022-10-27 11:34:17 +02:00
Irevoire	0bbf80186f	push the snapshot files	2022-10-27 11:34:17 +02:00
Irevoire	b6a0abea9f	fix the index deletion when the index doesn’t exists but would be created by one of the autobatched tasks	2022-10-27 11:34:16 +02:00
Irevoire	5303bbffab	fix the last rule about merging the allow_index_creation	2022-10-27 11:34:16 +02:00
Irevoire	fc944c39a5	simplify the code A LOT and create less false positive	2022-10-27 11:34:16 +02:00
Irevoire	a1d4cc673d	add a whole new batch of tests around the index already exists / allow_index_creation	2022-10-27 11:34:16 +02:00
Irevoire	28d9f2c041	fix all the snapshot tests	2022-10-27 11:34:16 +02:00
Irevoire	d9218578e3	it probably works but it's also horrendous	2022-10-27 11:34:16 +02:00
Loïc Lecrenier	d20b5ddda0	Don't return an error when swapping 0 indexes	2022-10-27 11:34:16 +02:00
Loïc Lecrenier	11fee30f47	Apply review suggestions and stop using rtxn.commit	2022-10-27 11:34:15 +02:00
Loïc Lecrenier	17cd2a4aa0	Implement POST /indexes-swap	2022-10-27 11:34:15 +02:00
Loïc Lecrenier	28bd8b6c6b	Remove key from index_tasks database when the value is empty	2022-10-27 11:34:15 +02:00
Loïc Lecrenier	169f386418	Add some documentation to the index scheduler	2022-10-27 11:34:15 +02:00
Irevoire	66c3b93ef1	fix all the snapshot tests in the dump	2022-10-27 11:34:15 +02:00
Loïc Lecrenier	bdb17954d2	Fix bug where assert used != instead of == And update snapshot tests.	2022-10-27 11:34:15 +02:00
Loïc Lecrenier	23b01a58df	cargo fmt	2022-10-27 11:34:14 +02:00
Loïc Lecrenier	ec3391808d	Fix date parsing for task queries Use rfc3339 or YYYY-MM-DD. Add a day to the parsed date when it is an excluded lower bound and the YYYY-MM-DD was used. Also the Query type does not need to be serialisable anymore	2022-10-27 11:34:14 +02:00
Loïc Lecrenier	10a547df4f	Apply suggestions from code review Co-authored-by: Clément Renault <clement@meilisearch.com> Apply suggestions from code review Co-authored-by: Clément Renault <clement@meilisearch.com> Co-authored-by: Tamo <tamo@meilisearch.com> Apply suggestions from code review Co-authored-by: Clément Renault <clement@meilisearch.com> Co-authored-by: Tamo <tamo@meilisearch.com> Apply code review suggestion Co-authored-by: Clément Renault <clement@meilisearch.com>	2022-10-27 11:34:14 +02:00
Loïc Lecrenier	22cf0559fe	Implement task date filters before/after enqueued/started/finished at	2022-10-27 11:34:14 +02:00
Irevoire	5765883600	fix the auto-generated details	2022-10-27 11:34:14 +02:00
Tamo	cff003c928	remove the unused variants from the autobatcher	2022-10-27 11:34:14 +02:00
Tamo	ab8f1c2865	fix a bunch of snapshot tests	2022-10-27 11:34:13 +02:00
Tamo	6730e190db	fix the dumps tests since we added informations in the DumpTask	2022-10-27 11:34:13 +02:00
Kerollmops	50b8b9df6a	Delete the tasks content file once the transaction has been successfully committed	2022-10-27 11:34:13 +02:00
Kerollmops	ec0a5a9f01	Remove the useless r#union thing	2022-10-27 11:34:13 +02:00
Kerollmops	6460b78e08	Clean up the delete_persisted_task_data function	2022-10-27 11:34:13 +02:00
Kerollmops	d21651c968	Throw the error if we can't register the tasks in the store	2022-10-27 11:34:13 +02:00
Kerollmops	6e904d0997	Introduce a ProcessingTasks constructor	2022-10-27 11:34:12 +02:00
Kerollmops	b373d19831	Extract the must_stop flag out of the RwLock	2022-10-27 11:34:12 +02:00
Kerollmops	3cbfacb616	Prefer using an u64 instead of a usize in some places	2022-10-27 11:34:12 +02:00
Kerollmops	79c4275bfc	Delete the persisted data when we cancel a task	2022-10-27 11:34:12 +02:00
Kerollmops	f9c8fe5eaa	Use a tokio block_in_place method for potentially blocking tasks	2022-10-27 11:34:12 +02:00
Kerollmops	c2ec4a089b	Put the original URL query in the tasks details	2022-10-27 11:34:12 +02:00
Kerollmops	751e9bac3b	Add the tasks cancel route to cancel tasks	2022-10-27 11:34:11 +02:00
Kerollmops	290945e258	Update the canceledBy and finishedAt fields	2022-10-27 11:34:11 +02:00
Kerollmops	725158b454	Introduce the core algorithm of task cancelation	2022-10-27 11:34:11 +02:00
Kerollmops	b2c5bc67b7	Add more enum-iterator related stuff	2022-10-27 11:34:11 +02:00
Kerollmops	591527a99d	Prefer using TaskDeletion in the dumps	2022-10-27 11:34:11 +02:00
Kerollmops	1ca9a67c49	Introduce the task cancelation task type	2022-10-27 11:34:11 +02:00
Kerollmops	f177c97671	Add the canceled task status	2022-10-27 11:34:10 +02:00
Kerollmops	703ba7a1fb	Introduce the ProcessingTasks struct	2022-10-27 11:34:10 +02:00
Kerollmops	c9523c6f39	Use the indexation-abortion milli's branch	2022-10-27 11:34:10 +02:00
Kerollmops	e645c4c4d6	Remove the meilisearch-auth milli dependency	2022-10-27 11:34:10 +02:00
Loïc Lecrenier	ea60d35c71	Delete a task's persisted data when appropriate	2022-10-27 11:34:10 +02:00
Tamo	f7e546eea3	make the tests compile again	2022-10-27 11:34:10 +02:00
Tamo	b45c430165	fix the analytics	2022-10-27 11:34:10 +02:00
Tamo	634eb52926	extract the create_app function for the tests	2022-10-27 11:34:09 +02:00
Tamo	d1a6fb2971	bump enum-iter and fix a bunch of error messages	2022-10-27 11:34:09 +02:00
Tamo	bea81ae37b	fix meilisearch-http	2022-10-27 11:34:09 +02:00
Tamo	9e85f050b2	fix the tests	2022-10-27 11:34:09 +02:00
Tamo	2f748480a1	share the rtxn between the access to the tasks and to the indexes	2022-10-27 11:34:09 +02:00
Tamo	6bd6321226	dump the content of the dump tasks instead of recreating at import time with wrong API keys	2022-10-27 11:34:08 +02:00
Tamo	655705eb2b	remove useless todo	2022-10-27 11:34:08 +02:00
Tamo	9fe24fbff2	get rids of the useless Seek before creating a grenad reader	2022-10-27 11:34:08 +02:00
Tamo	83f3c5ec57	flush the dump-writer only once everything has been inserted	2022-10-27 11:34:08 +02:00
Tamo	78ce29f461	apply most style comments of the review	2022-10-27 11:34:08 +02:00
Tamo	dd70daaae3	Update dump/src/error.rs Co-authored-by: Clément Renault <clement@meilisearch.com>	2022-10-27 11:34:08 +02:00
Tamo	d0e91555d1	rebase on index-scheduler	2022-10-27 11:34:08 +02:00
Tamo	e0221fc0a3	fix a synchronization bug while importing tasks	2022-10-27 11:34:07 +02:00
Tamo	a9eeb070b8	fix all the errors code and settings issues when importing a dump v2	2022-10-27 11:34:07 +02:00
Tamo	3872a1b8d1	fix all the error codes	2022-10-27 11:34:07 +02:00
Tamo	ba150f2127	commit after creating an index	2022-10-27 11:34:07 +02:00
Tamo	554600dfd8	fix the deletion of the data.ms in case of errors	2022-10-27 11:34:07 +02:00
Tamo	e9295c03ce	the index-scheduler needs to wake-up after importing a dump	2022-10-27 11:34:06 +02:00
Tamo	955d3339f0	remove the dbg	2022-10-27 11:34:06 +02:00
Tamo	d481669b7e	fix the content_file import	2022-10-27 11:34:06 +02:00
Tamo	dd506e5d87	stop dumping the current dumping task as enqueued so it's not looping for ever	2022-10-27 11:34:06 +02:00
Tamo	208c785697	add a bufwriter on the documents	2022-10-27 11:34:06 +02:00
Tamo	d976e680c5	first mostly working version	2022-10-27 11:34:06 +02:00
Tamo	c051166bcc	update the API a little bit	2022-10-27 11:34:05 +02:00
Tamo	72a906ae75	fix the tests	2022-10-27 11:34:05 +02:00
Tamo	b7f9c94f4a	write the dump export	2022-10-27 11:34:05 +02:00
Loïc Lecrenier	8954b1bd1d	Fix number of deleted tasks details after duplicate task deletion	2022-10-27 11:34:05 +02:00
Loïc Lecrenier	8defad6c38	Add task deletion tests where the same task is deleted twice	2022-10-27 11:34:05 +02:00
Loïc Lecrenier	f32b973945	Return an error when calling DELETE /tasks with an empty query	2022-10-27 11:34:04 +02:00
Loïc Lecrenier	fbd2be2ec8	Apply suggested changes from PR review	2022-10-27 11:34:04 +02:00
Loïc Lecrenier	441417447e	Avoid creating two read txn at the same time	2022-10-27 11:34:04 +02:00
Loïc Lecrenier	8c6aeaada5	Update snapshot tests following git rebase that fixes a bug	2022-10-27 11:34:04 +02:00
Loïc Lecrenier	8bb0fcd144	Finish first draft of the DELETE /tasks route	2022-10-27 11:34:04 +02:00
Loïc Lecrenier	9522b75454	Continue implementation of task deletion 1. Matched tasks are a roaring bitmap 2. Start implementation in meilisearch-http 3. Snapshots use meili-snap 4. Rename to TaskDeletion	2022-10-27 11:34:03 +02:00
Kerollmops	e4d461ecba	Make sure that we do not batch tasks from different indexes	2022-10-27 11:34:03 +02:00
Kerollmops	b029369653	Add a test to check different indexes autobatching	2022-10-27 11:34:03 +02:00
Kerollmops	408d00136c	Extract index creation rights and simplify the autobatcher rules	2022-10-27 11:34:03 +02:00
Kerollmops	2c24c7d403	Fix invalid import of tasks types	2022-10-27 11:34:03 +02:00
Tamo	7034803712	move the API key in meilisearch_types	2022-10-27 11:34:02 +02:00
Tamo	c192146fbe	remove an unused file	2022-10-27 11:34:02 +02:00
Tamo	b6c84e53ba	uncomment a task serialization test	2022-10-27 11:34:02 +02:00
Tamo	2f1eb78b1d	refactor the Task a little bit	2022-10-27 11:34:02 +02:00
Tamo	510ce9fc51	start moving a lot of task types to meilisearch_types	2022-10-27 11:34:01 +02:00
Tamo	fa4c1de019	store md5 instead of the whole snapshots	2022-10-27 11:34:01 +02:00
Loïc Lecrenier	3e4337c91f	Add meili-snap crate to make writing snapshot tests easier	2022-10-27 11:34:01 +02:00
Tamo	0af00f6b32	fix all the import and comment most of the dump v6	2022-10-27 11:34:01 +02:00
Tamo	141a1c9464	push the document_format and settings I forgot in the previous PR	2022-10-27 11:34:00 +02:00
Tamo	667c282e19	get rids of the index crate + the document_types crate	2022-10-27 11:34:00 +02:00
Loïc Lecrenier	9a74ea0943	Fix compiler errors related autobatching option of the index scheduler	2022-10-27 11:34:00 +02:00
Loïc Lecrenier	eabac9676b	Fix typo and remove useless code in tests	2022-10-27 11:34:00 +02:00
Loïc Lecrenier	ab4e649221	Apply suggestions from code review Co-authored-by: Tamo <tamo@meilisearch.com>	2022-10-27 11:34:00 +02:00
Loïc Lecrenier	568199fc0d	Add more task deletion tests	2022-10-27 11:33:59 +02:00
Loïc Lecrenier	13a72f8757	Use more complete snapshot tests for the index scheduler	2022-10-27 11:33:59 +02:00
Loïc Lecrenier	4c55c30027	Add a DetailsView type and improve index scheduler snapshots The DetailsView type is necessary because serde incorrectly deserialises the `Details` type, so the database fails to correctly decode Tasks	2022-10-27 11:33:59 +02:00
Loïc Lecrenier	dc81992eb2	Implement TaskDeletion in the index scheduler	2022-10-27 11:33:59 +02:00
Kerollmops	fe84f2648b	Allow a user to disable the auto batching system	2022-10-27 11:33:59 +02:00
Kerollmops	e2a766acb5	Add a test to check that it works without autobatching	2022-10-27 11:33:58 +02:00
Kerollmops	db9d1b18ca	Remove the IndexScheduler::notify method	2022-10-27 11:33:58 +02:00
Kerollmops	19c6f8303f	Make sure that the index-scheduler tick loop is rerun after processing	2022-10-27 11:33:58 +02:00
Kerollmops	b311eb3bed	Add a test that verifies that sending multiple tasks works	2022-10-27 11:33:58 +02:00
Tamo	f026ac3115	remove unused files	2022-10-27 11:33:58 +02:00
Tamo	f176382b34	fix the tests	2022-10-27 11:33:57 +02:00
Tamo	4bd9e4d723	write a bunch of tests that goes through the whole compat layers	2022-10-27 11:33:57 +02:00
Tamo	c6f4fb5f7d	remove the warnings	2022-10-27 11:33:57 +02:00
Tamo	2ae0806773	rewrite the update file API	2022-10-27 11:33:57 +02:00
Tamo	7579a363ab	finish the dump reader API, the dump Writer API now needs to be updated	2022-10-27 11:33:57 +02:00
Tamo	0284764b5e	start dumping the update files to a known format	2022-10-27 11:33:56 +02:00
Tamo	9117fde712	fix the compat between v3 and v4	2022-10-27 11:33:56 +02:00
Tamo	f622ef9836	remove the ununsed snapshot files	2022-10-27 11:33:56 +02:00
Tamo	dc0f307d61	remove all warnings	2022-10-27 11:33:56 +02:00
Tamo	06fadb3004	write the compat layer from v2 to v3	2022-10-27 11:33:55 +02:00
Tamo	6107540ad4	remove old compat files	2022-10-27 11:33:55 +02:00
Tamo	7e18f92635	write the dump v2 import	2022-10-27 11:33:55 +02:00
Tamo	43496b97bd	make the open function public	2022-10-27 11:33:55 +02:00
Tamo	6f327a00c7	fix some warnings	2022-10-27 11:33:55 +02:00
Tamo	58ef80a2a7	rebase on main	2022-10-27 11:33:54 +02:00
Tamo	22ffbf3676	write and test the compat layer from v3 to v4	2022-10-27 11:33:54 +02:00
Tamo	089106a970	write and test the dump v3 import	2022-10-27 11:33:54 +02:00
Tamo	026f6fb06a	fix the test once again	2022-10-27 11:33:54 +02:00
Tamo	efe0a5f422	finish the test for the compatibility between v4 and v5	2022-10-27 11:33:53 +02:00
Tamo	47e0288747	rewrite the compat API to something more generic	2022-10-27 11:33:53 +02:00
Tamo	2f47443458	rename a few things for consistency	2022-10-27 11:33:53 +02:00
Tamo	a8128678a4	implement the dump v4 import	2022-10-27 11:33:53 +02:00
Tamo	c50b44039e	add the compat layer between v5 and v6	2022-10-27 11:33:53 +02:00
Tamo	6dcc5851b5	get rids of the trait in most places	2022-10-27 11:33:52 +02:00
Tamo	0972587cfc	start writting the compat layer between v5 and v6	2022-10-27 11:33:52 +02:00
Tamo	afd5fe0783	test the dump v5	2022-10-27 11:33:52 +02:00
Tamo	1473a71e33	write the v5 dump import	2022-10-27 11:33:52 +02:00
Tamo	101f55ce8b	introduce the index metadata	2022-10-27 11:33:52 +02:00
Tamo	e845cc2b6f	fix the tests	2022-10-27 11:33:51 +02:00
Tamo	7bd6f63001	implement the dump reader v6	2022-10-27 11:33:51 +02:00
Tamo	699ae1b190	start implementing a skeleton of the v1 dump reader	2022-10-27 11:33:51 +02:00
Tamo	f041d474a5	move the DumpWriter and Error to their own module	2022-10-27 11:33:51 +02:00
Tamo	ece6c3f6e7	fix the dump export	2022-10-27 11:33:51 +02:00
Tamo	87a6a337aa	write a dump exporter	2022-10-27 11:33:51 +02:00
Clément Renault	123f47dbc4	Create the index only if the task has the rights to do so	2022-10-27 11:33:50 +02:00
Clément Renault	068a4b2884	Correctly batch tasks with different index creation rights	2022-10-27 11:33:50 +02:00
Clément Renault	87212cfd20	Use a ControlFlow in the autobatcher function	2022-10-27 11:33:50 +02:00
Kerollmops	f1b1cfdbcc	IndexDeletion operation have ClearAll details	2022-10-27 11:33:50 +02:00
Kerollmops	a083c9e452	Only mark the first clear document with the amount of cleared documents	2022-10-27 11:33:50 +02:00
Kerollmops	b24b13b036	Let the tick function set the Failed status itself	2022-10-27 11:33:50 +02:00
Kerollmops	566c15fb74	Fill an IndexDeletion task with the number of documents removed	2022-10-27 11:33:49 +02:00
Kerollmops	6b3b05fb73	Panic if we encountered a wring KindWithContent type	2022-10-27 11:33:49 +02:00
Kerollmops	36e5efde0d	Update the tasks statuses	2022-10-27 11:33:49 +02:00
Kerollmops	2fbdd104b8	Implement the IndexDeletion batch operation	2022-10-27 11:33:49 +02:00
Kerollmops	da363a92ac	Implement the IndexUpdate batch operation	2022-10-27 11:33:49 +02:00
Kerollmops	0543cba6eb	Implement the IndexCreate batch operation	2022-10-27 11:33:48 +02:00
Kerollmops	cf6084151b	Make sure that meilisearch-http works without index wrapper	2022-10-27 11:33:48 +02:00
Kerollmops	c70f375669	Implement ErrorCode on the heed Error	2022-10-27 11:33:48 +02:00
Kerollmops	91e13c2824	Implement ErrorCode on the milli::Error type	2022-10-27 11:33:48 +02:00
Kerollmops	d76634a36c	Remove the Index wrapper and use milli::Index directly	2022-10-27 11:33:48 +02:00
Kerollmops	9e8242c57d	Remove the IndexRename operation	2022-10-27 11:33:48 +02:00
Kerollmops	5fa214abb1	Move the IndexScheduler to the root of the index-scheduler crate	2022-10-27 11:33:47 +02:00
Kerollmops	9a9e98fb77	Add a TODO about the index creation	2022-10-27 11:33:47 +02:00
Kerollmops	5d21c790ef	Make clippy happy	2022-10-27 11:33:47 +02:00
Kerollmops	31de33d5ee	Implement a recursive indexation for the index-related operations	2022-10-27 11:33:47 +02:00
Kerollmops	3b343a930e	Fix meilisearch-http to use the new DocumentImport batch operation	2022-10-27 11:33:47 +02:00
Kerollmops	07286fcc79	Implement the SettingsAndDocumentImport batch operation	2022-10-27 11:33:47 +02:00
Kerollmops	f68906f5dc	Merge both DocumentAddition/Update into one DocumentImport variant	2022-10-27 11:33:46 +02:00
Kerollmops	5174c78f87	Implement the DocumentClear batch operation	2022-10-27 11:33:46 +02:00
Kerollmops	025bb5f616	Implement the DocumentClearAndSettings batch operation	2022-10-27 11:33:46 +02:00
Kerollmops	41ec737e73	Implement the Settings batch operation	2022-10-27 11:33:46 +02:00
Kerollmops	7b4a913704	Implement the DocumentUpdate batch operation	2022-10-27 11:33:46 +02:00
Kerollmops	a6a1043abb	Implement the DocumentDeletion batch operation	2022-10-27 11:33:46 +02:00
Tamo	7a0f17c912	remove an old unworking part of the batch execution	2022-10-27 11:33:45 +02:00
Tamo	c2899fe9b2	bring back the IndexMeta and IndexStats in meilisearch-http	2022-10-27 11:33:45 +02:00
Tamo	c759fd6924	fix import bug	2022-10-27 11:33:45 +02:00
Tamo	fba9aa214a	remove the create_app macro	2022-10-27 11:33:45 +02:00
Tamo	2c8f1a43e9	get rids of meilisearch-lib	2022-10-27 11:33:44 +02:00
Tamo	0ba1c46e19	fix a deadlock	2022-10-27 11:33:44 +02:00
Tamo	22bfb5a7a0	remove `Clone` from the `IndexScheduler`	2022-10-27 11:33:44 +02:00
Tamo	d8d3499aec	remove a bunch of comments	2022-10-27 11:33:44 +02:00
Tamo	64e132ce53	move as many fields as possible out of the IndexScheduler	2022-10-27 11:33:44 +02:00
Tamo	9e1f38ec7c	move the test function in the test module	2022-10-27 11:33:44 +02:00
Tamo	6f4dcc0c38	start implementing some logic to test the internal states of the scheduler	2022-10-27 11:33:43 +02:00
Tamo	84cd5cef0b	fix the tests	2022-10-27 11:33:43 +02:00
Tamo	ae86a8ccd6	slightly refactor the autobatching tests	2022-10-27 11:33:43 +02:00
Tamo	ce2dfecc03	connect the new scheduler to meilisearch-http officially. I can index documents and do search	2022-10-27 11:33:43 +02:00
Tamo	cb4feabca2	implements the get_tasks	2022-10-27 11:33:43 +02:00
Tamo	19154e48fe	fix all compilation errors	2022-10-27 11:33:42 +02:00
Irevoire	8d51c1f389	wip integrating the scheduler in meilisearch-http	2022-10-27 11:33:42 +02:00
Irevoire	250410495c	start integrating the index-scheduler in meilisearch-lib	2022-10-27 11:33:42 +02:00
Irevoire	8f0fd35358	add insta::json for later	2022-10-27 11:33:42 +02:00
Irevoire	8770e07397	I can index documents without meilisearch	2022-10-27 11:33:42 +02:00
Tamo	edd8344dc9	wip	2022-10-27 11:33:42 +02:00
Tamo	e547552702	create the end Batch type for all Index* operations	2022-10-27 11:33:41 +02:00
Tamo	925971809a	create the end Batch type for all Document* operation	2022-10-27 11:33:41 +02:00
Tamo	1ea9c0b4c0	write most of the run loop	2022-10-27 11:33:41 +02:00
Tamo	4846a7c501	use faux in the file-store	2022-10-27 11:33:41 +02:00
Tamo	9ff0fe952e	split the run function in two	2022-10-27 11:33:41 +02:00
Tamo	a8b18b2c96	fix the register test	2022-10-27 11:33:40 +02:00
Tamo	5436b996ab	reduce the size of the snapshots	2022-10-27 11:33:40 +02:00
Tamo	7d0c8a3379	test the register tasks	2022-10-27 11:33:40 +02:00
Tamo	fc098022c7	start integrating the index-scheduler in the meilisearch codebase	2022-10-27 11:33:40 +02:00
Tamo	b816535e33	greatly reduce the number of warnings	2022-10-27 11:33:40 +02:00
Tamo	38e4ffe73c	fix smol typo	2022-10-27 11:33:40 +02:00
Tamo	366a344474	get rids of the horrendous spinlock in favor of synchronoise	2022-10-27 11:33:39 +02:00
Tamo	7b6673dc1d	implement the index swap in the index mapper	2022-10-27 11:33:39 +02:00
Tamo	03aca2e452	move the index mapping logic in another structure	2022-10-27 11:33:39 +02:00
Tamo	4129783019	migrate the index handling code in a different file + implements the create index	2022-10-27 11:33:39 +02:00
Tamo	1804416afa	reintroduce the uuid mapping for the indexes	2022-10-27 11:33:39 +02:00
Tamo	c97d51a624	add a bunch of tests	2022-10-27 11:33:39 +02:00
Tamo	803f2157af	split the DocumentAdditionOrUpdate in two tasks; DocumentAddition and DocumentUpdate	2022-10-27 11:33:38 +02:00
Tamo	b7c5b71a53	starts importing the real tasks	2022-10-27 11:33:38 +02:00
Tamo	5cc8f96237	get rids of the auto-generated mains	2022-10-27 11:33:38 +02:00
Tamo	94e29a9f5f	extract the index abstraction out of the index-scheduler in its own module	2022-10-27 11:33:38 +02:00
Tamo	48138c21a9	rename the update-file-store to file-store since it can store any kind of file	2022-10-27 11:33:38 +02:00
Tamo	76597fc382	import the update_file_store in the index-scheduler	2022-10-27 11:33:37 +02:00
Tamo	2afb381f95	get rids of nelson	2022-10-27 11:33:37 +02:00
Tamo	a9844bd4f6	move the update file store to another crate with as little dependencies as possible	2022-10-27 11:33:37 +02:00
Tamo	a0588d6b94	finishes the global skelton of the auto-batcher	2022-10-27 11:33:37 +02:00
Tamo	b3c9b128d9	polish the global structure of the batch creation	2022-10-27 11:33:37 +02:00
Irevoire	448f44f631	move the autobatcher logic to another file	2022-10-27 11:33:36 +02:00
Tamo	f638774764	add the document format file	2022-10-27 11:33:36 +02:00
Tamo	516860f342	fix the create_new_batch method	2022-10-27 11:33:36 +02:00
Tamo	6b9689a1c0	fix the whole batchKind thingy	2022-10-27 11:33:36 +02:00
Tamo	af0f5d6c0c	implements most operations	2022-10-27 11:33:36 +02:00
Tamo	5a7fcf2688	fix a few typos	2022-10-27 11:33:35 +02:00
Tamo	30d2b24689	implements the index deletion, creation and swap	2022-10-27 11:33:35 +02:00
Tamo	72b2e68de4	makes the updates getters smoother to uses	2022-10-27 11:33:35 +02:00
Tamo	7879189c6b	make the project compile again	2022-10-27 11:33:35 +02:00
Tamo	46b8ebcab4	fix the file store	2022-10-27 11:33:35 +02:00
Tamo	fa742f60e8	make the file store entirely synchronous, including the file deletion	2022-10-27 11:33:35 +02:00
Tamo	a7aa92df5f	fix most of the index module	2022-10-27 11:33:34 +02:00
Irevoire	d8b8e04ad1	wip porting the index back in the scheduler	2022-10-27 11:33:34 +02:00
Irevoire	fe330e1be9	add a little bit of documentation	2022-10-27 11:33:34 +02:00
Tamo	2c4e5ce8be	implements the filter query	2022-10-27 11:33:34 +02:00
Tamo	705af94fd7	add the task to the index db in the register task	2022-10-27 11:33:34 +02:00
Tamo	ed745591e1	split the scheduler into multiples files	2022-10-27 11:33:34 +02:00
Tamo	22d24dba56	implement the get_batch method	2022-10-27 11:33:33 +02:00
Tamo	1a47949063	START THE REWRITE OF THE INDEX SCHEDULER: index & register has been implemented	2022-10-27 11:33:33 +02:00
bors[bot]	ab1800551f	Merge #2922 2922: Add new error when using /keys without masterkey set r=ManyTheFish a=vishalsodani # Pull Request ## Related issue Fixes #2918 Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Co-authored-by: vishalsodani <vishalsodani@rediffmail.com>	2022-10-27 09:13:11 +00:00
vishalsodani	689bef7ad2	fmt the code	2022-10-27 14:09:38 +05:30
vishalsodani	89c40c83c3	refactor code to avoid cloning	2022-10-27 14:08:29 +05:30
vishalsodani	03ba830ab2	uncomment tests	2022-10-27 12:59:28 +05:30
vishalsodani	9cf3ff72a3	fix checking of master key as per review comment	2022-10-27 12:56:18 +05:30
Samyak S Sarnayak	d35afa0cf5	Change consecutive phrase search grouping logic Co-authored-by: ManyTheFish <many@meilisearch.com>	2022-10-26 23:10:48 +05:30
Samyak S Sarnayak	752d031010	Update phrase search to use new `execute` method	2022-10-26 23:07:20 +05:30
bors[bot]	25ec51e783	Merge #2601 2601: Ease search result pagination r=Kerollmops a=ManyTheFish # Summary This PR is a prototype enhancing search results pagination (#2577) # Todo - [x] Update the API to return the number of pages and allow users to directly choose a page instead of computing an offset - [x] Change computation of `total_pages` in order to have an exact count - [x] compute query tree exhaustively - [x] compute distinct exhaustively # Small Documentation ## Default search query request: ```sh curl \ -X POST 'http://localhost:7700/indexes/movies/search' \ -H 'Content-Type: application/json' \ --data-binary '{ "q": "botman" }' ``` result: ```json { "hits":[...], "query":"botman", "processingTimeMs":5, "hitsPerPage":20, "page":1, "totalPages":4, "totalHits":66 } ``` ## Search query with offset parameter request: ```sh curl \ -X POST 'http://localhost:7700/indexes/movies/search' \ -H 'Content-Type: application/json' \ --data-binary '{ "q": "botman", "offset": 0 }' ``` result: ```json { "hits":[...], "query":"botman", "processingTimeMs":3, "limit":20, "offset":0, "estimatedTotalHits":66 } ``` ## Search query selecting page with page parameter request: ```sh curl \ -X POST 'http://localhost:7700/indexes/movies/search' \ -H 'Content-Type: application/json' \ --data-binary '{ "q": "botman", "page": 2 }' ``` result: ```json { "hits":[...], "query":"botman", "processingTimeMs":5, "hitsPerPage":20, "page":2, "totalPages":4, "totalHits":66 } ``` # Related fixes #2577 ## In charge of the feature Core: `@ManyTheFish` Docs: `@guimachiavelli` Integration: `@bidoubiwa` Co-authored-by: ManyTheFish <many@meilisearch.com>	2022-10-26 16:10:58 +00:00
ManyTheFish	f4021273b8	Add is_finite_pagination method to SearchQuery	2022-10-26 18:08:29 +02:00
unvalley	c7322f704c	Fix cargo clippy errors Dont apply clippy for tests for now Fix clippy warnings of filter-parser package parent 8352febd646ec4bcf56a44161e5c4dce0e55111f author unvalley <38400669+unvalley@users.noreply.github.com> 1666325847 +0900 committer unvalley <kirohi.code@gmail.com> 1666791316 +0900 Update .github/workflows/rust.yml Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com> Allow clippy lint too_many_argments Allow clippy lint needless_collect Allow clippy lint too_many_arguments and type_complexity Fix for clippy warnings comparison_chains Fix for clippy warnings vec_init_then_push Allow clippy lint should_implement_trait Allow clippy lint drop_non_drop Fix lifetime clipy warnings in filter-paprser Execute cargo fmt Fix clippy remaining warnings Fix clippy remaining warnings again and allow lint on each place	2022-10-27 01:04:23 +09:00
unvalley	811f156031	Execute cargo clippy --fix	2022-10-27 01:00:00 +09:00
unvalley	d8fed1f7a9	Add clippy job Add Run Clippy to bors.toml	2022-10-27 01:00:00 +09:00
bors[bot]	2e539249cb	Merge #619 619: Refactor the Facets databases to enable incremental indexing r=curquiza a=loiclec # Pull Request ## What does this PR do? Party fixes https://github.com/meilisearch/milli/issues/605 by making the indexing of the facet databases (i.e. `facet_id_f64_docids` and `facet_id_string_docids`) incremental. It also closes #327 and https://github.com/meilisearch/meilisearch/issues/2820 . Two more untracked bugs were also fixed: 1. The facet distribution algorithm did not respect the `maxFacetValues` parameter when there were only a few candidate document ids. 2. The structure of the levels > 0 of the facet databases were not updated following the deletion of documents ## How to review this PR First, read this comment to get an overview of the changes. Then, based on this comment, raise any concerns you might have about: 1. the new structure of the databases 2. the algorithms for sort, facet distribution, and range search 3. the new/removed heed codecs Then, weigh in on the following concerns: 1. adding `fuzzcheck` as a fuzz-only dependency may add too much complexity for the benefits it provides 2. the `ByteSliceRef` and `StrRefCodec` are misnamed or should not exist 3. the new behaviour of facet distributions can be considered incorrect 4. incremental deletion is useless given that documents are always deleted in bulk ## What's left for me to do 1. Re-read everything once to make sure I haven't forgotten anything 2. Wait for the results of the benchmarks and see if (1) they provide enough information (2) there was any change in performance, especially for search queries. Then, maybe, spend some time optimising the code. 3. Test whether the `info`/`http-ui` crates survived the refactor ## Old structure of the `facet_id_f64_docids` and `facet_id_string_docids` databases Previously, these two databases had different but conceptually similar structures. For each field id, the facet number database had the following format: ``` ┌───────────────────────────────┬───────────────────────────────┬───────────────┐ ┌───────┐ │ 1.2 – 2 │ 3.4 – 100 │ 102 – 104 │ │Level 2│ │ │ │ │ └───────┘ │ a, b, d, f, z │ c, d, e, f, g │ u, y │ ├───────────────┬───────────────┼───────────────┬───────────────┼───────────────┤ ┌───────┐ │ 1.2 – 1.3 │ 1.6 – 2 │ 3.4 – 12 │ 12.3 – 100 │ 102 – 104 │ │Level 1│ │ │ │ │ │ │ └───────┘ │ a, b, d, z │ a, b, f │ c, d, g │ e, f │ u, y │ ├───────┬───────┼───────┬───────┼───────┬───────┼───────┬───────┼───────┬───────┤ ┌───────┐ │ 1.2 │ 1.3 │ 1.6 │ 2 │ 3.4 │ 12 │ 12.3 │ 100 │ 102 │ 104 │ │Level 0│ │ │ │ │ │ │ │ │ │ │ │ └───────┘ │ a, b │ d, z │ b, f │ a, f │ c, d │ g │ e │ e, f │ y │ u │ └───────┴───────┴───────┴───────┴───────┴───────┴───────┴───────┴───────┴───────┘ ``` where the first line is the key of the database, consisting of : - the field id - the level height - the left and right bound of the group and the second line is the value of the database, consisting of: - a bitmap of all the docids that have a facet value within the bounds The `facet_id_string_docids` had a similar structure: ``` ┌───────────────────────────────┬───────────────────────────────┬───────────────┐ ┌───────┐ │ 0 – 3 │ 4 – 7 │ 8 – 9 │ │Level 2│ │ │ │ │ └───────┘ │ a, b, d, f, z │ c, d, e, f, g │ u, y │ ├───────────────┬───────────────┼───────────────┬───────────────┼───────────────┤ ┌───────┐ │ 0 – 1 │ 2 – 3 │ 4 – 5 │ 6 – 7 │ 8 – 9 │ │Level 1│ │ "ab" – "ac" │ "ba" – "bac" │ "gaf" – "gal" │"form" – "wow" │ "woz" – "zz" │ └───────┘ │ a, b, d, z │ a, b, f │ c, d, g │ e, f │ u, y │ ├───────┬───────┼───────┬───────┼───────┬───────┼───────┬───────┼───────┬───────┤ ┌───────┐ │ "ab" │ "ac" │ "ba" │ "bac" │ "gaf" │ "gal" │ "form"│ "wow" │ "woz" │ "zz" │ │Level 0│ │ "AB" │ " Ac" │ "ba " │ "Bac" │ " GAF"│ "gal" │ "Form"│ " wow"│ "woz" │ "ZZ" │ └───────┘ │ a, b │ d, z │ b, f │ a, f │ c, d │ g │ e │ e, f │ y │ u │ └───────┴───────┴───────┴───────┴───────┴───────┴───────┴───────┴───────┴───────┘ ``` where, at level 0, the key is: * the normalised facet value (string) and the value is: * the original facet value (string) * a bitmap of all the docids that have this normalised string facet value At level 1, the key is: * the left bound of the range as an index in level 0 * the right bound of the range as an index in level 0 and the value is: * the left bound of the range as a normalised string * the right bound of the range as a normalised string * a bitmap of all the docids that have a string facet value within the bounds At level > 1, the key is: * the left bound of the range as an index in level 0 * the right bound of the range as an index in level 0 and the value is: * a bitmap of all the docids that have a string facet value within the bounds ## New structure of the `facet_id_f64_docids` and `facet_id_string_docids` databases Now both the `facet_id_f64_docids` and `facet_id_string_docids` databases have the exact same structure: ``` ┌───────────────────────────────┬───────────────────────────────┬───────────────┐ ┌───────┐ │ "ab" (2) │ "gaf" (2) │ "woz" (1) │ │Level 2│ │ │ │ │ └───────┘ │ [a, b, d, f, z] │ [c, d, e, f, g] │ [u, y] │ ├───────────────┬───────────────┼───────────────┬───────────────┼───────────────┤ ┌───────┐ │ "ab" (2) │ "ba" (2) │ "gaf" (2) │ "form" (2) │ "woz" (2) │ │Level 1│ │ │ │ │ │ │ └───────┘ │ [a, b, d, z] │ [a, b, f] │ [c, d, g] │ [e, f] │ [u, y] │ ├───────┬───────┼───────┬───────┼───────┬───────┼───────┬───────┼───────┬───────┤ ┌───────┐ │ "ab" │ "ac" │ "ba" │ "bac" │ "gaf" │ "gal" │ "form"│ "wow" │ "woz" │ "zz" │ │Level 0│ │ │ │ │ │ │ │ │ │ │ │ └───────┘ │ [a, b]│ [d, z]│ [b, f]│ [a, f]│ [c, d]│ [g] │ [e] │ [e, f]│ [y] │ [u] │ └───────┴───────┴───────┴───────┴───────┴───────┴───────┴───────┴───────┴───────┘ ``` where for all levels, the key is a `FacetGroupKey<T>` containing: * the field id (`u16`) * the level height (`u8`) * the left bound of the range (`T`) and the value is a `FacetGroupValue` containing: * the number of elements from the level below that are part of the range (`u8`, =0 for level 0) * a bitmap of all the docids that have a facet value within the bounds (`RoaringBitmap`) The right bound of the range is now implicit, it is equal to `Excluded(next_left_bound)`. In the code, the key is always encoded using `FacetGroupKeyCodec<C>` where `C` is the codec used to encode the facet value (either `OrderedF64Codec` or `StrRefCodec`) and the value is encoded with `FacetGroupValueCodec`. Since both databases share the same structure, we can implement almost all operations only once by treating the facet value as a byte slice (i.e. `FacetGroupKey<&[u8]>` encoded as `FacetGroupKeyCodec<ByteSliceRef>`). This is, in my opinion, a big simplification. The reason for changing the structure of the databases is to make it possible to incrementally add a facet value to an existing database. Since the `facet_id_string_docids` used to store indices to `level 0` in all levels > 0, adding an element to level 0 would potentially invalidate all the indices. Note that the original string value of a facet is no longer stored in this database. ## Incrementally adding a facet value Here I describe how we can add a facet value to the new database incrementally. If we want to add the document with id `z` and facet value `gap`., then we want to add/modify the elements highlighted below in pink: <img width="946" alt="Screenshot 2022-09-12 at 10 14 54" src="https://user-images.githubusercontent.com/6040237/189605532-fe4b0f52-e13d-4b3c-92d9-10c705953e3d.png"> which results in: <img width="662" alt="Screenshot 2022-09-12 at 10 23 29" src="https://user-images.githubusercontent.com/6040237/189607015-c3a37588-b825-43c2-878a-f8f85c000b94.png"> * one element was added in level 0 * one key/value was modified in level 1 * one value was modified in level 2 Adding this element was easy since we could simply add it to level 0 and then increase the `group_size` part of the value for the level above. However, in order to keep the structure balanced, we can't always do this. If the group size reaches a threshold (`max_group_size`), then we split the node into two. For example, let's imagine that `max_group_size` is `4` and we add the docid `y` with facet value `gas`. First, we add it in level 0: <img width="904" alt="Screenshot 2022-09-12 at 10 30 40" src="https://user-images.githubusercontent.com/6040237/189608391-531f9df1-3424-4f1f-8344-73eb194570e5.png"> Then, we realise that the group size of its parent is going to reach the maximum group size (=4) and thus we split the parent into two nodes: <img width="919" alt="Screenshot 2022-09-12 at 10 33 16" src="https://user-images.githubusercontent.com/6040237/189608884-66f87635-1fc6-41d2-a459-87c995491ac4.png"> and since we inserted an element in level 1, we also update level 2 accordingly, by increasing the group size of the parent: <img width="915" alt="Screenshot 2022-09-12 at 10 34 42" src="https://user-images.githubusercontent.com/6040237/189609233-d4a893ff-254a-48a7-a5ad-c0dc337f23ca.png"> We also have two other parameters: * `group_size` is the default group size when building the database from scratch * `min_level_size` is the minimum number of elements that a level should contain When the highest level size is greater than `group_size * min_level_size`, then we create an additional level above it. There is one more edge case for the insertion algorithm. While we normally don't modify the existing left bounds of a key, we have to do it if the facet value being inserted is smaller than the first left bound. For example, inserting `"aa"` with the docid `w` would change the database to: <img width="756" alt="Screenshot 2022-09-12 at 10 41 56" src="https://user-images.githubusercontent.com/6040237/189610637-a043ef71-7159-4bf1-b4fd-9903134fc095.png"> The root of the code for incremental indexing is the `FacetUpdateIncremental` builder. ## Incrementally removing a facet value TODO: the algorithm was implemented and works, but its current API is: `fn delete(self, facet_value, single_docid)`. It removes the given document id from all keys containing the given facet value. I don't think it is the right way to implement it anymore. Perhaps a bitmap of docids should be given instead. This is fairly easy to do. But since we batch document deletions together (because of soft deletion), it's not clear to me anymore that incremental deletion should be implemented at all. ## Bulk insertion While it's faster to incrementally add a single facet value to the database, it is sometimes slower to repeatedly add facet values one-by-one instead of doing it in bulk. For example, during initial indexing, we'd like to build the database from a list of facet values and associated document ids in one go. The `FacetUpdateBulk` builder provides a way to do so. It works by: 1. clearing all levels > 0 from the DB 2. adding all new elements in level 0 3. rebuilding the higher levels from scratch The algorithm for bulk insertion is the same as the previous one. ## Choosing between incremental and bulk insertion On my computer, I measured that is about 50x slower to add N facet values incrementally than it is to re-build a database with N facet values in level 0. Therefore, we dynamically choose to use either incremental insertion or bulk insertion based on (1) the number of existing elements in level 0 of the database and (2) the number of facet values from the new documents. This is imprecise but is mainly aimed at avoiding the worst-case scenario where the incremental insertion method is used repeatedly millions of times. ## Fuzz-testing Potentially controversial: I fuzz-tested incremental addition and deletion using fuzzcheck, which found many bugs. The fuzz-test consists of inserting/deleting facet values and docids in succession, each operation is processed with different parameters for `group_size`, `max_group_size`, and `min_level_size`. After all the operations are processed, the content of level 0 is compared to the content of an equivalent structure with a simple and easily-checked implementation. Furthermore, we check that the database has a correct structure (all groups from levels > 0 correctly combine the content of their children). I also visualised the code coverage found by the fuzz-test. It covered 100% of the relevant code except for `unreachable/panic` statements and errors returned by `heed`. The fuzz-test and the fuzzcheck dependency are only compiled when `cargo fuzzcheck` is used. For now, the dependency is from a local path on my computer, but it can be changed to a crate version if we decide to keep it. ## Algorithms operating on the facet databases There are four important algorithms making use of the facet databases: 1. Sort, ascending 2. Sort, descending 3. Facet distribution 4. Range search Previously, the implementation of all four algorithms was based on a number of iterators specific to each database kind (number or string): `FacetNumberRange`, `FacetNumberRevRange`, `FacetNumberIter` (with a reversed and reducing/non-reducing option), `FacetStringGroupRange`, `FacetStringGroupRevRange`, `FacetStringLevel0Range`, `FacetStringLevel0RevRange`, and `FacetStringIter` (reversed + reducing/non-reducing). Now, all four algorithms have a unique implementation shared by both the string and number databases. There are four functions: 1. `ascending_facet_sort` in `search/facet/facet_sort_ascending.rs` 2. `descending_facet_sort` in `search/facet/facet_sort_descending.rs` 3. `iterate_over_facet_distribution` in `search/facet/facet_distribution_iter.rs` 4. `find_docids_of_facet_within_bounds` in `search/facet/facet_range_search.rs` I have tried to test them with some snapshot tests but more testing could still be done. I don't think that the performance of these algorithms regressed, but that will need to be confirmed by benchmarks. ## Change of behaviour for facet distributions Previously, the original string value of a facet was stored in the level 0 of `facet_id_string_docids `. This is no longer the case. The original string value was used in the implementation of the facet distribution algorithm. Now, to recover it, we pick a random document id which contains the normalised string value and look up the original one in `field_id_docid_facet_strings`. As a consequence, it may be that the string value returned in the field distribution does not appear in any of the candidates. For example, ```json { "id": 0, "colour": "RED" } { "id": 1, "colour": "red" } ``` Facet distribution for the `colour` field among the candidates `[1]`: ``` { "RED": 1 } ``` Here, "RED" was given as the original facet value even though it does not appear in the document id `1`. ## Heed codecs A number of heed codecs related to the facet databases were removed: * `FacetLevelValueF64Codec` * `FacetLevelValueU32Codec` * `FacetStringLevelZeroCodec` * `StringValueCodec` * `FacetStringZeroBoundsValueCodec` * `FacetValueStringCodec` * `FieldDocIdFacetStringCodec` * `FieldDocIdFacetF64Codec` They were replaced by: * `FacetGroupKeyCodec<C>` (replaces all key codecs for the facet databases) * `FacetGroupValueCodec` (replaces all value codecs for the facet databases) * `FieldDocIdFacetCodec<C>` (replaces `FieldDocIdFacetStringCodec` and `FieldDocIdFacetF64Codec`) Since the associated encoded item of `FacetGroupKeyCodec<C>` is `FacetKey<T>` and we often work with `FacetKey<&[u8]>` and `FacetKey<&str>`, then we need to have codecs that encode values of type `&str` and `&[u8]`. The existing `ByteSlice` and `Str` codecs do not work for that purpose (their `EItem` are `[u8]` and `str`), I have also created two new codecs: * `ByteSliceRef` is a codec with a `EItem = DItem = &[u8]` * `StrRefCodec` is a codec with a `EItem = DItem = &str` I have also factored out the code used to encode an ordered f64 into its own `OrderedF64Codec`. Co-authored-by: Loïc Lecrenier <loic@meilisearch.com>	2022-10-26 15:04:53 +00:00
Samyak S Sarnayak	488d31ecdf	Run cargo fmt	2022-10-26 19:09:45 +05:30
Samyak S Sarnayak	af33d22f25	Consecutive is false when at least 1 stop word is surrounded by words	2022-10-26 19:09:45 +05:30
Samyak S Sarnayak	f1da623af3	Add test for phrase search with stop words and all criteria at once Moved the actual test into a separate function used by both the existing test and the new test.	2022-10-26 19:09:44 +05:30
Samyak S Sarnayak	77f1ff019b	Simplify stop word checking in create_primitive_query	2022-10-26 19:09:44 +05:30
Samyak S Sarnayak	2aa11afb87	Fix panic when phrase contains only one stop word and nothing else	2022-10-26 19:09:42 +05:30
Samyak S Sarnayak	bb9ce3c5c5	Run cargo fmt	2022-10-26 19:09:03 +05:30
Samyak S Sarnayak	d187b32a28	Fix snapshots to use new phrase type	2022-10-26 19:09:03 +05:30
Samyak S Sarnayak	c8c666c6a6	Use resolve_phrase in exactness and typo criteria	2022-10-26 19:09:01 +05:30
Samyak S Sarnayak	3e190503e6	Search for closest non-stop words in proximity criteria	2022-10-26 19:08:34 +05:30
Samyak S Sarnayak	709ab3c14c	Increment position even when it's a stop word in exactness criteria	2022-10-26 19:08:33 +05:30
Samyak S Sarnayak	ef13c6a5b6	Perform filter after enumerate to keep origin indices	2022-10-26 19:08:33 +05:30
Samyak S Sarnayak	6a10b679ca	Add test for phrase search with stop words Originally written by ManyTheFish here: https://gist.github.com/ManyTheFish/f840e37cb2d2e029ce05396b4d540762 Co-authored-by: ManyTheFish <many@meilisearch.com>	2022-10-26 19:08:32 +05:30
Samyak S Sarnayak	62816dddde	[WIP] Fix phrase search containing stop words Fixes #661 and meilisearch/meilisearch#2905	2022-10-26 19:08:06 +05:30
Loïc Lecrenier	54c0cf93fe	Merge remote-tracking branch 'origin/main' into facet-levels-refactor	2022-10-26 15:13:34 +02:00
bors[bot]	365f44c39b	Merge #668 668: Fix many Clippy errors part 2 r=ManyTheFish a=ehiggs This brings us a step closer to enforcing clippy on each build. # Pull Request ## Related issue This does not fix any issue outright, but it is a second round of fixes for clippy after https://github.com/meilisearch/milli/pull/665. This should contribute to fixing https://github.com/meilisearch/milli/pull/659. ## What does this PR do? Satisfies many issues for clippy. The complaints are mostly: * Passing reference where a variable is already a reference. * Using clone where a struct already implements `Copy` * Using `ok_or_else` when it is a closure that returns a value instead of using the closure to call function (hence we use `ok_or`) * Unambiguous lifetimes don't need names, so we can just use `'_` * Using `return` when it is not needed as we are on the last expression of a function. ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Ewan Higgs <ewan.higgs@gmail.com>	2022-10-26 12:16:24 +00:00
Loïc Lecrenier	2fa85a24ec	Remove outdated files from http-ui/ and infos/ ... that were reintroduced after a rebase	2022-10-26 14:09:35 +02:00
Loïc Lecrenier	631e9910da	Depend on released version of fuzzcheck from crates.io	2022-10-26 14:06:59 +02:00
Loïc Lecrenier	2741756248	Merge remote-tracking branch 'origin/main' into facet-levels-refactor	2022-10-26 14:03:23 +02:00
bors[bot]	d3f95e6c69	Merge #671 671: Update version for the next release (v0.35.0) in Cargo.toml files r=Kerollmops a=meili-bot ⚠️ This PR is automatically generated. Check the new version is the expected one before merging. Co-authored-by: curquiza <curquiza@users.noreply.github.com>	2022-10-26 11:58:05 +00:00
Loïc Lecrenier	b7f2428961	Fix formatting and warning after rebasing from main	2022-10-26 13:49:33 +02:00
Loïc Lecrenier	3b1f908e5e	Revert behaviour of facet distribution to what it was before Where the docid that is used to get the original facet string value definitely belongs to the candidates	2022-10-26 13:48:01 +02:00
Loïc Lecrenier	14ca8048a8	Add some documentation on how to run the facet db fuzzer	2022-10-26 13:48:01 +02:00
Loïc Lecrenier	206a3e00e5	cargo fmt	2022-10-26 13:48:01 +02:00
Loïc Lecrenier	f198b20c42	Add facet deletion tests that use both the incremental and bulk methods + update deletion snapshots to the new database format	2022-10-26 13:47:46 +02:00
Loïc Lecrenier	e3ba1fc883	Make deletion tests for both soft-deletion and hard-deletion	2022-10-26 13:47:46 +02:00
Loïc Lecrenier	ab5e56fd16	Add document deletion snapshot tests and tests for hard-deletion	2022-10-26 13:47:46 +02:00
Loïc Lecrenier	d885de1600	Add option to avoid soft deletion of documents	2022-10-26 13:47:46 +02:00
Loïc Lecrenier	ee1abfd1c1	Ignore files generated by fuzzcheck	2022-10-26 13:47:46 +02:00
Loïc Lecrenier	2295e0e3ce	Use real delete function in facet indexing fuzz tests By deleting multiple docids at once instead of one-by-one	2022-10-26 13:47:46 +02:00
Loïc Lecrenier	acc8caebe6	Add link to GitHub PR to document of update/facet module	2022-10-26 13:47:46 +02:00
Loïc Lecrenier	a034a1e628	Move StrRefCodec and ByteSliceRefCodec to their own files	2022-10-26 13:47:46 +02:00
Loïc Lecrenier	1165ba2171	Make facet deletion incremental	2022-10-26 13:47:04 +02:00
Loïc Lecrenier	0ade699873	Don't crash when failing to decode using StrRef codec	2022-10-26 13:47:04 +02:00
Loïc Lecrenier	d0109627b9	Fix a bug in facet_range_search and add documentation	2022-10-26 13:47:04 +02:00
Loïc Lecrenier	a2270b7432	Change fuzzcheck dependency to point to git repository	2022-10-26 13:47:04 +02:00
Loïc Lecrenier	1ecd3bb822	Fix bug in FieldDocIdFacetCodec	2022-10-26 13:47:04 +02:00
Loïc Lecrenier	51961e1064	Polish some details	2022-10-26 13:47:04 +02:00
Loïc Lecrenier	cb8442a119	Further unify facet databases of f64s and strings	2022-10-26 13:47:04 +02:00
Loïc Lecrenier	3baa34d842	Fix compiler errors/warnings	2022-10-26 13:47:04 +02:00
Loïc Lecrenier	86d9f50b9c	Fix bugs in incremental facet indexing with variable parameters e.g. add one facet value incrementally with a group_size = X and then add another one with group_size = Y It is not actually possible to do so with the public API of milli, but I wanted to make sure the algorithm worked well in those cases anyway. The bugs were found by fuzzing the code with fuzzcheck, which I've added to milli as a conditional dev-dependency. But it can be removed later.	2022-10-26 13:47:04 +02:00
Loïc Lecrenier	de52a9bf75	Improve documentation of some facet-related algorithms	2022-10-26 13:47:04 +02:00
Loïc Lecrenier	985a94adfc	cargo fmt	2022-10-26 13:47:04 +02:00
Loïc Lecrenier	b1ab09196c	Remove outdated TODOs	2022-10-26 13:47:04 +02:00
Loïc Lecrenier	3d7ed3263f	Fix bug in string facet distribution with few candidates	2022-10-26 13:47:04 +02:00
Loïc Lecrenier	fca4577e23	Return original string in facet distributions, work on facet tests	2022-10-26 13:47:04 +02:00
Loïc Lecrenier	27454e9828	Document and refine facet indexing algorithms	2022-10-26 13:47:04 +02:00
Loïc Lecrenier	bee3c23b45	Add comparison benchmark between bulk and incremental facet indexing	2022-10-26 13:47:04 +02:00
Loïc Lecrenier	b2f01ad204	Refactor facet database tests	2022-10-26 13:47:04 +02:00
Loïc Lecrenier	9026867d17	Give same interface to bulk and incremental facet indexing types + cargo fmt, oops, sorry for the bad history :(	2022-10-26 13:47:04 +02:00
Loïc Lecrenier	330c9eb1b2	Rename facet codecs and refine FacetsUpdate API	2022-10-26 13:47:04 +02:00
Loïc Lecrenier	485a72306d	Refactor facet-related codecs	2022-10-26 13:47:04 +02:00
Loïc Lecrenier	9b55e582cd	Add FacetsUpdate type that wraps incremental and bulk indexing methods	2022-10-26 13:47:04 +02:00
Loïc Lecrenier	3d145d7f48	Merge the two <facetttype>_faceted_documents_ids methods into one	2022-10-26 13:47:04 +02:00
Loïc Lecrenier	982efab88f	Fix encoding bugs in facet databases	2022-10-26 13:47:04 +02:00
Loïc Lecrenier	079ed4a992	Add more snapshots	2022-10-26 13:47:04 +02:00
Loïc Lecrenier	afdf87f6f7	Fix bugs in asc/desc criterion and facet indexing	2022-10-26 13:47:04 +02:00
Loïc Lecrenier	a7201ece04	cargo fmt	2022-10-26 13:47:04 +02:00
Loïc Lecrenier	36296bbb20	Add facet incremental indexing snapshot tests + fix bug	2022-10-26 13:47:04 +02:00
Loïc Lecrenier	07ff92c663	Add more snapshots from facet tests	2022-10-26 13:47:04 +02:00
Loïc Lecrenier	61252248fb	Fix some facet indexing bugs	2022-10-26 13:47:04 +02:00
Loïc Lecrenier	68cbcdf08b	Fix compile errors/warnings in http-ui and infos	2022-10-26 13:47:04 +02:00
Loïc Lecrenier	85824ee203	Try to make facet indexing incremental	2022-10-26 13:47:04 +02:00
Loïc Lecrenier	d30c89e345	Fix compile error+warnings in new tests	2022-10-26 13:46:46 +02:00
Loïc Lecrenier	e8a156d682	Reorganise facets database indexing code	2022-10-26 13:46:46 +02:00
Loïc Lecrenier	fb8d23deb3	Reintroduce db_snap! for facet databases	2022-10-26 13:46:14 +02:00
Loïc Lecrenier	e570c23153	Reintroduce asc/desc functionality	2022-10-26 13:46:14 +02:00
Loïc Lecrenier	bd2c0e1ab6	Remove unused code	2022-10-26 13:46:14 +02:00
Loïc Lecrenier	39a4a0a362	Reintroduce filter range search and facet extractors	2022-10-26 13:46:14 +02:00
Loïc Lecrenier	22d80eeaf9	Reintroduce facet deletion functionality	2022-10-26 13:46:14 +02:00
Loïc Lecrenier	6cc91824c1	Remove unused heed codec files	2022-10-26 13:46:14 +02:00
Loïc Lecrenier	5a904cf29d	Reintroduce facet distribution functionality	2022-10-26 13:46:14 +02:00
Loïc Lecrenier	b8a1caad5e	Add range search and incremental indexing algorithm	2022-10-26 13:46:14 +02:00
Loïc Lecrenier	63ef0aba18	Start porting facet distribution and sort to new database structure	2022-10-26 13:46:14 +02:00
Loïc Lecrenier	7913d6365c	Update Facets indexing to be compatible with new database structure	2022-10-26 13:46:14 +02:00
Loïc Lecrenier	c3f49f766d	Prepare refactor of facets database Prepare refactor of facets database	2022-10-26 13:46:14 +02:00
curquiza	e883bccc76	Update version for the next release (v0.35.0) in Cargo.toml files	2022-10-26 11:43:54 +00:00
bors[bot]	c8f16530d5	Merge #616 616: Introduce an indexation abortion function when indexing documents r=Kerollmops a=Kerollmops Co-authored-by: Kerollmops <clement@meilisearch.com> Co-authored-by: Clément Renault <clement@meilisearch.com>	2022-10-26 11:41:18 +00:00
Ewan Higgs	9d27ac8a2e	Ignore too many arguments to functions.	2022-10-25 21:22:53 +02:00
Ewan Higgs	42cdc38c7b	Allow weird ranges like 1..=0 to pass clippy. Everything else is just a warning and exit code will be 0.	2022-10-25 21:12:59 +02:00
Ewan Higgs	2ce025a906	Fixes after rebase to fix new issues.	2022-10-25 20:58:31 +02:00
Ewan Higgs	17f7922bfc	Remove unneeded lifetimes.	2022-10-25 20:49:04 +02:00
Ewan Higgs	6b2fe94192	Fixes for clippy bringing us down to 18 remaining issues. This brings us a step closer to enforcing clippy on each build.	2022-10-25 20:49:02 +02:00
vishalsodani	f0ecacb58d	add implementation for no master key set and fix tests	2022-10-25 22:41:48 +05:30
ManyTheFish	68c9751d49	Fix clippy	2022-10-25 16:08:07 +02:00
bors[bot]	004c09a8e2	Merge #669 669: Add method to create a new Index with specific creation dates r=irevoire a=loiclec This functionality is needed to implement the import of dumps correctly. Co-authored-by: Loïc Lecrenier <loic.lecrenier@me.com>	2022-10-25 12:44:43 +00:00
Loïc Lecrenier	36bd66281d	Add method to create a new Index with specific creation dates	2022-10-25 14:37:56 +02:00
bors[bot]	9aef1031ca	Merge #2961 2961: Changed error message for config file r=curquiza a=LunarMarathon # Pull Request ## Related issue Fixes #2959 ## What does this PR do? - Changed the error message as required in the issue - changed config to configuration ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: LunarMarathon <lmaytan24@gmail.com>	2022-10-25 11:36:47 +00:00
LunarMarathon	bc2a161f62	Change err mess for config	2022-10-25 16:16:34 +05:30
bors[bot]	d11a6e187f	Merge #639 639: Reduce the size of the word_pair_proximity database r=loiclec a=loiclec # Pull Request ## What does this PR do? Fixes #634 Now, the value corresponding to the key `prox word1 word2` in the `word_pair_proximity_docids` database contains the ids of the documents in which: - `word1` is followed by `word2` - the minimum number of words between `word1` and `word2` is `prox-1` Before this PR, the `word_pair_proximity_docids` had keys with the format `word1 word2 prox` and the value contained the ids of the documents in which either: - `word1` is followed by `word2` after a minimum of `prox-1` words in between them - `word2` is followed by `word1` after a minimum of `prox-2` words As a consequence of this change, calls such as: ``` let docids = word_pair_proximity_docids.get(rtxn, (word1, word2, prox)); ``` have to be replaced with: ``` let docids1 = word_pair_proximity_docids.get(rtxn, (prox, word1, word2)) ; let docids2 = word_pair_proximity_docids.get(rtxn, (prox-1, word2, word1)) ; let docids = docids1 \| docids2; ``` ## Phrase search The PR also fixes two bugs in the `resolve_phrase` function. The first bug is that a phrase containing twice the same word would always return zero documents (e.g. `"dog eats dog"`). The second bug occurs with a phrase such as "fox is smarter than a dog"` and the document with the text: ``` fox or dog? a fox is smarter than a dog ``` In that case, the phrase search would not return the documents because: * we only have the key `fox dog 2` in `word_pair_proximity_docids` * but the implementation of `resolve_phrase` looks for `fox dog 5`, which returns 0 documents ### New implementation of `resolve_phrase` Given the phrase: ``` fox is smarter than a dog ``` We select the document ids corresponding to all of the following keys in `word_pair_proximity_docids`: - `1 fox is` - `1 is smarter` - `1 smarter than` - (etc.) - `1 fox smarter` OR `2 fox smarter` - `1 is than` OR `2 is than` - ... - `1 than dog` OR `2 than dog` ## Benchmark Results Indexing: ``` group indexing_main_d94339a8 indexing_word-pair-proximity-docids-refactor_2983dd8e ----- ---------------------- ----------------------------------------------------- indexing/-geo-delete-facetedNumber-facetedGeo-searchable- 1.19 40.7±11.28ms ? ?/sec 1.00 34.3±4.16ms ? ?/sec indexing/-movies-delete-facetedString-facetedNumber-searchable- 1.62 11.3±3.77ms ? ?/sec 1.00 7.0±1.56ms ? ?/sec indexing/-movies-delete-facetedString-facetedNumber-searchable-nested- 1.00 12.5±2.62ms ? ?/sec 1.07 13.4±4.24ms ? ?/sec indexing/-songs-delete-facetedString-facetedNumber-searchable- 1.26 50.2±12.63ms ? ?/sec 1.00 39.8±20.25ms ? ?/sec indexing/-wiki-delete-searchable- 1.83 269.1±16.11ms ? ?/sec 1.00 146.8±6.12ms ? ?/sec indexing/Indexing geo_point 1.00 47.2±0.46s ? ?/sec 1.00 47.3±0.56s ? ?/sec indexing/Indexing movies in three batches 1.42 12.7±0.13s ? ?/sec 1.00 9.0±0.07s ? ?/sec indexing/Indexing movies with default settings 1.40 10.2±0.07s ? ?/sec 1.00 7.3±0.06s ? ?/sec indexing/Indexing nested movies with default settings 1.22 7.8±0.11s ? ?/sec 1.00 6.4±0.13s ? ?/sec indexing/Indexing nested movies without any facets 1.24 7.3±0.07s ? ?/sec 1.00 5.9±0.06s ? ?/sec indexing/Indexing songs in three batches with default settings 1.14 47.6±0.67s ? ?/sec 1.00 41.8±0.63s ? ?/sec indexing/Indexing songs with default settings 1.13 44.1±0.74s ? ?/sec 1.00 38.9±0.76s ? ?/sec indexing/Indexing songs without any facets 1.19 42.0±0.66s ? ?/sec 1.00 35.2±0.48s ? ?/sec indexing/Indexing songs without faceted numbers 1.20 44.3±1.40s ? ?/sec 1.00 37.0±0.48s ? ?/sec indexing/Indexing wiki 1.39 862.9±9.95s ? ?/sec 1.00 622.6±27.11s ? ?/sec indexing/Indexing wiki in three batches 1.40 934.4±5.97s ? ?/sec 1.00 665.7±4.72s ? ?/sec indexing/Reindexing geo_point 1.01 15.9±0.39s ? ?/sec 1.00 15.7±0.28s ? ?/sec indexing/Reindexing movies with default settings 1.15 288.8±25.03ms ? ?/sec 1.00 250.4±2.23ms ? ?/sec indexing/Reindexing songs with default settings 1.01 4.1±0.06s ? ?/sec 1.00 4.1±0.03s ? ?/sec indexing/Reindexing wiki 1.41 1484.7±20.59s ? ?/sec 1.00 1052.0±19.89s ? ?/sec ``` Search Wiki: <details> <pre> group search_wiki_main_d94339a8 search_wiki_word-pair-proximity-docids-refactor_2983dd8e ----- ------------------------- -------------------------------------------------------- smol-wiki-articles.csv: basic placeholder/ 1.02 25.8±0.21µs ? ?/sec 1.00 25.4±0.19µs ? ?/sec smol-wiki-articles.csv: basic with quote/"film" 1.00 441.7±2.57µs ? ?/sec 1.00 442.3±2.41µs ? ?/sec smol-wiki-articles.csv: basic with quote/"france" 1.00 357.0±2.63µs ? ?/sec 1.00 358.3±2.65µs ? ?/sec smol-wiki-articles.csv: basic with quote/"japan" 1.00 239.4±2.24µs ? ?/sec 1.00 240.2±1.82µs ? ?/sec smol-wiki-articles.csv: basic with quote/"machine" 1.00 180.3±2.40µs ? ?/sec 1.00 180.0±1.08µs ? ?/sec smol-wiki-articles.csv: basic with quote/"miles" "davis" 1.00 9.1±0.03ms ? ?/sec 1.03 9.3±0.04ms ? ?/sec smol-wiki-articles.csv: basic with quote/"mingus" 1.00 3.6±0.01ms ? ?/sec 1.03 3.7±0.02ms ? ?/sec smol-wiki-articles.csv: basic with quote/"rock" "and" "roll" 1.00 34.0±0.11ms ? ?/sec 1.03 35.1±0.13ms ? ?/sec smol-wiki-articles.csv: basic with quote/"spain" 1.00 162.0±0.88µs ? ?/sec 1.00 161.9±0.98µs ? ?/sec smol-wiki-articles.csv: basic without quote/film 1.01 164.4±1.46µs ? ?/sec 1.00 163.1±1.58µs ? ?/sec smol-wiki-articles.csv: basic without quote/france 1.00 1698.3±7.37µs ? ?/sec 1.00 1697.7±11.53µs ? ?/sec smol-wiki-articles.csv: basic without quote/japan 1.00 1154.0±23.61µs ? ?/sec 1.00 1150.7±9.27µs ? ?/sec smol-wiki-articles.csv: basic without quote/machine 1.00 524.6±3.45µs ? ?/sec 1.01 528.1±4.56µs ? ?/sec smol-wiki-articles.csv: basic without quote/miles davis 1.00 13.5±0.05ms ? ?/sec 1.02 13.8±0.05ms ? ?/sec smol-wiki-articles.csv: basic without quote/mingus 1.00 4.1±0.02ms ? ?/sec 1.03 4.2±0.01ms ? ?/sec smol-wiki-articles.csv: basic without quote/rock and roll 1.00 49.0±0.19ms ? ?/sec 1.03 50.4±0.22ms ? ?/sec smol-wiki-articles.csv: basic without quote/spain 1.00 412.2±3.35µs ? ?/sec 1.00 412.9±2.81µs ? ?/sec smol-wiki-articles.csv: prefix search/c 1.00 383.9±2.53µs ? ?/sec 1.00 383.4±2.44µs ? ?/sec smol-wiki-articles.csv: prefix search/g 1.00 433.4±2.53µs ? ?/sec 1.00 432.8±2.52µs ? ?/sec smol-wiki-articles.csv: prefix search/j 1.00 424.3±2.05µs ? ?/sec 1.00 424.0±2.15µs ? ?/sec smol-wiki-articles.csv: prefix search/q 1.00 154.0±1.93µs ? ?/sec 1.00 153.5±1.04µs ? ?/sec smol-wiki-articles.csv: prefix search/t 1.04 658.5±91.93µs ? ?/sec 1.00 631.4±3.89µs ? ?/sec smol-wiki-articles.csv: prefix search/x 1.00 446.2±2.09µs ? ?/sec 1.00 445.6±3.13µs ? ?/sec smol-wiki-articles.csv: proximity/april paris 1.02 3.4±0.39ms ? ?/sec 1.00 3.3±0.01ms ? ?/sec smol-wiki-articles.csv: proximity/diesel engine 1.00 1022.1±17.52µs ? ?/sec 1.00 1017.7±8.16µs ? ?/sec smol-wiki-articles.csv: proximity/herald sings 1.01 1872.5±97.70µs ? ?/sec 1.00 1862.2±8.57µs ? ?/sec smol-wiki-articles.csv: proximity/tea two 1.00 295.2±34.91µs ? ?/sec 1.00 296.6±4.08µs ? ?/sec smol-wiki-articles.csv: typo/Disnaylande 1.00 3.4±0.51ms ? ?/sec 1.04 3.5±0.01ms ? ?/sec smol-wiki-articles.csv: typo/aritmetric 1.00 3.6±0.01ms ? ?/sec 1.00 3.7±0.01ms ? ?/sec smol-wiki-articles.csv: typo/linax 1.00 167.5±1.28µs ? ?/sec 1.00 167.1±2.65µs ? ?/sec smol-wiki-articles.csv: typo/migrosoft 1.01 217.9±1.84µs ? ?/sec 1.00 216.2±1.61µs ? ?/sec smol-wiki-articles.csv: typo/nympalidea 1.00 2.9±0.01ms ? ?/sec 1.10 3.1±0.01ms ? ?/sec smol-wiki-articles.csv: typo/phytogropher 1.00 3.0±0.23ms ? ?/sec 1.08 3.3±0.01ms ? ?/sec smol-wiki-articles.csv: typo/sisan 1.00 234.6±1.38µs ? ?/sec 1.01 235.8±1.67µs ? ?/sec smol-wiki-articles.csv: typo/the fronce 1.00 104.4±0.84µs ? ?/sec 1.00 103.9±0.81µs ? ?/sec smol-wiki-articles.csv: words/Abraham machin 1.02 675.5±4.74µs ? ?/sec 1.00 662.1±5.13µs ? ?/sec smol-wiki-articles.csv: words/Idaho Bellevue pizza 1.02 1004.5±11.07µs ? ?/sec 1.00 989.5±13.08µs ? ?/sec smol-wiki-articles.csv: words/Kameya Tokujirō mingus monk 1.00 1650.8±10.92µs ? ?/sec 1.00 1643.2±10.77µs ? ?/sec smol-wiki-articles.csv: words/Ulrich Hensel meilisearch milli 1.00 5.4±0.03ms ? ?/sec 1.00 5.4±0.02ms ? ?/sec smol-wiki-articles.csv: words/the black saint and the sinner lady and the good doggo 1.00 32.9±0.10ms ? ?/sec 1.00 32.8±0.10ms ? ?/sec </pre> </details> Search songs: <details> <pre> group search_songs_main_d94339a8 search_songs_word-pair-proximity-docids-refactor_2983dd8e ----- -------------------------- --------------------------------------------------------- smol-songs.csv: asc + default/Notstandskomitee 1.00 3.0±0.01ms ? ?/sec 1.01 3.0±0.04ms ? ?/sec smol-songs.csv: asc + default/charles 1.00 2.2±0.01ms ? ?/sec 1.01 2.2±0.01ms ? ?/sec smol-songs.csv: asc + default/charles mingus 1.00 3.1±0.01ms ? ?/sec 1.01 3.1±0.01ms ? ?/sec smol-songs.csv: asc + default/david 1.00 2.9±0.01ms ? ?/sec 1.00 2.9±0.01ms ? ?/sec smol-songs.csv: asc + default/david bowie 1.00 4.5±0.02ms ? ?/sec 1.00 4.5±0.02ms ? ?/sec smol-songs.csv: asc + default/john 1.00 3.1±0.01ms ? ?/sec 1.01 3.2±0.01ms ? ?/sec smol-songs.csv: asc + default/marcus miller 1.00 5.0±0.02ms ? ?/sec 1.00 5.0±0.02ms ? ?/sec smol-songs.csv: asc + default/michael jackson 1.00 4.7±0.02ms ? ?/sec 1.00 4.7±0.02ms ? ?/sec smol-songs.csv: asc + default/tamo 1.00 1463.4±12.17µs ? ?/sec 1.01 1481.5±8.83µs ? ?/sec smol-songs.csv: asc + default/thelonious monk 1.00 4.4±0.01ms ? ?/sec 1.00 4.4±0.02ms ? ?/sec smol-songs.csv: asc/Notstandskomitee 1.01 2.6±0.01ms ? ?/sec 1.00 2.6±0.01ms ? ?/sec smol-songs.csv: asc/charles 1.00 473.6±3.70µs ? ?/sec 1.01 476.8±22.17µs ? ?/sec smol-songs.csv: asc/charles mingus 1.01 780.1±3.90µs ? ?/sec 1.00 773.6±4.60µs ? ?/sec smol-songs.csv: asc/david 1.00 757.6±4.50µs ? ?/sec 1.00 760.7±5.20µs ? ?/sec smol-songs.csv: asc/david bowie 1.00 1131.2±8.68µs ? ?/sec 1.00 1130.7±8.36µs ? ?/sec smol-songs.csv: asc/john 1.00 668.9±6.48µs ? ?/sec 1.00 669.9±2.78µs ? ?/sec smol-songs.csv: asc/marcus miller 1.00 959.8±7.10µs ? ?/sec 1.00 958.9±4.72µs ? ?/sec smol-songs.csv: asc/michael jackson 1.01 1076.7±16.73µs ? ?/sec 1.00 1070.8±7.34µs ? ?/sec smol-songs.csv: asc/tamo 1.00 70.4±0.55µs ? ?/sec 1.00 70.5±0.51µs ? ?/sec smol-songs.csv: asc/thelonious monk 1.01 2.9±0.01ms ? ?/sec 1.00 2.9±0.01ms ? ?/sec smol-songs.csv: basic filter: <=/Notstandskomitee 1.00 162.0±0.91µs ? ?/sec 1.01 163.6±1.72µs ? ?/sec smol-songs.csv: basic filter: <=/charles 1.00 38.3±0.24µs ? ?/sec 1.01 38.7±0.31µs ? ?/sec smol-songs.csv: basic filter: <=/charles mingus 1.01 85.3±0.44µs ? ?/sec 1.00 84.6±0.47µs ? ?/sec smol-songs.csv: basic filter: <=/david 1.01 32.4±0.25µs ? ?/sec 1.00 32.1±0.24µs ? ?/sec smol-songs.csv: basic filter: <=/david bowie 1.00 68.6±0.99µs ? ?/sec 1.01 68.9±0.88µs ? ?/sec smol-songs.csv: basic filter: <=/john 1.04 26.1±0.37µs ? ?/sec 1.00 25.1±0.22µs ? ?/sec smol-songs.csv: basic filter: <=/marcus miller 1.00 76.7±0.39µs ? ?/sec 1.01 77.3±0.61µs ? ?/sec smol-songs.csv: basic filter: <=/michael jackson 1.00 95.5±0.66µs ? ?/sec 1.01 96.3±0.79µs ? ?/sec smol-songs.csv: basic filter: <=/tamo 1.03 26.2±0.36µs ? ?/sec 1.00 25.3±0.23µs ? ?/sec smol-songs.csv: basic filter: <=/thelonious monk 1.00 140.7±1.36µs ? ?/sec 1.01 142.7±0.88µs ? ?/sec smol-songs.csv: basic filter: TO/Notstandskomitee 1.00 165.4±1.25µs ? ?/sec 1.00 165.7±1.72µs ? ?/sec smol-songs.csv: basic filter: TO/charles 1.01 40.6±0.57µs ? ?/sec 1.00 40.1±0.54µs ? ?/sec smol-songs.csv: basic filter: TO/charles mingus 1.01 87.1±0.80µs ? ?/sec 1.00 86.3±0.61µs ? ?/sec smol-songs.csv: basic filter: TO/david 1.02 34.5±0.26µs ? ?/sec 1.00 33.7±0.24µs ? ?/sec smol-songs.csv: basic filter: TO/david bowie 1.00 70.6±0.38µs ? ?/sec 1.00 70.6±0.68µs ? ?/sec smol-songs.csv: basic filter: TO/john 1.02 27.5±0.77µs ? ?/sec 1.00 26.9±0.21µs ? ?/sec smol-songs.csv: basic filter: TO/marcus miller 1.01 79.8±0.76µs ? ?/sec 1.00 79.3±1.27µs ? ?/sec smol-songs.csv: basic filter: TO/michael jackson 1.00 98.3±0.54µs ? ?/sec 1.00 98.0±0.88µs ? ?/sec smol-songs.csv: basic filter: TO/tamo 1.03 27.9±0.23µs ? ?/sec 1.00 27.1±0.32µs ? ?/sec smol-songs.csv: basic filter: TO/thelonious monk 1.00 142.5±1.36µs ? ?/sec 1.02 145.2±0.98µs ? ?/sec smol-songs.csv: basic placeholder/ 1.00 49.4±0.34µs ? ?/sec 1.00 49.3±0.45µs ? ?/sec smol-songs.csv: basic with quote/"Notstandskomitee" 1.00 190.5±1.60µs ? ?/sec 1.01 191.8±2.10µs ? ?/sec smol-songs.csv: basic with quote/"charles" 1.00 165.0±1.13µs ? ?/sec 1.01 166.0±1.39µs ? ?/sec smol-songs.csv: basic with quote/"charles" "mingus" 1.00 1149.4±15.78µs ? ?/sec 1.02 1171.1±9.95µs ? ?/sec smol-songs.csv: basic with quote/"david" 1.00 236.5±1.61µs ? ?/sec 1.00 236.9±1.73µs ? ?/sec smol-songs.csv: basic with quote/"david" "bowie" 1.00 1384.8±9.02µs ? ?/sec 1.01 1393.8±11.39µs ? ?/sec smol-songs.csv: basic with quote/"john" 1.00 358.3±4.85µs ? ?/sec 1.00 358.9±1.75µs ? ?/sec smol-songs.csv: basic with quote/"marcus" "miller" 1.00 281.4±1.79µs ? ?/sec 1.01 285.6±3.24µs ? ?/sec smol-songs.csv: basic with quote/"michael" "jackson" 1.00 1328.4±8.01µs ? ?/sec 1.00 1334.6±8.00µs ? ?/sec smol-songs.csv: basic with quote/"tamo" 1.00 528.7±3.72µs ? ?/sec 1.01 533.4±5.31µs ? ?/sec smol-songs.csv: basic with quote/"thelonious" "monk" 1.00 1223.0±7.24µs ? ?/sec 1.02 1245.7±12.04µs ? ?/sec smol-songs.csv: basic without quote/Notstandskomitee 1.00 2.8±0.01ms ? ?/sec 1.00 2.8±0.01ms ? ?/sec smol-songs.csv: basic without quote/charles 1.00 273.3±2.06µs ? ?/sec 1.01 275.9±1.76µs ? ?/sec smol-songs.csv: basic without quote/charles mingus 1.00 2.3±0.01ms ? ?/sec 1.02 2.4±0.01ms ? ?/sec smol-songs.csv: basic without quote/david 1.00 434.3±3.86µs ? ?/sec 1.01 436.7±2.47µs ? ?/sec smol-songs.csv: basic without quote/david bowie 1.00 5.6±0.02ms ? ?/sec 1.01 5.7±0.02ms ? ?/sec smol-songs.csv: basic without quote/john 1.00 1322.5±9.98µs ? ?/sec 1.00 1321.2±17.40µs ? ?/sec smol-songs.csv: basic without quote/marcus miller 1.02 2.4±0.02ms ? ?/sec 1.00 2.4±0.01ms ? ?/sec smol-songs.csv: basic without quote/michael jackson 1.00 3.8±0.02ms ? ?/sec 1.01 3.9±0.01ms ? ?/sec smol-songs.csv: basic without quote/tamo 1.00 809.0±4.01µs ? ?/sec 1.01 819.0±6.22µs ? ?/sec smol-songs.csv: basic without quote/thelonious monk 1.00 3.8±0.02ms ? ?/sec 1.02 3.9±0.02ms ? ?/sec smol-songs.csv: big filter/Notstandskomitee 1.00 2.7±0.01ms ? ?/sec 1.01 2.8±0.01ms ? ?/sec smol-songs.csv: big filter/charles 1.00 266.5±1.34µs ? ?/sec 1.01 270.1±8.17µs ? ?/sec smol-songs.csv: big filter/charles mingus 1.00 651.0±5.40µs ? ?/sec 1.00 651.0±2.73µs ? ?/sec smol-songs.csv: big filter/david 1.00 1018.1±11.16µs ? ?/sec 1.00 1022.3±8.94µs ? ?/sec smol-songs.csv: big filter/david bowie 1.00 1912.2±11.13µs ? ?/sec 1.00 1919.8±8.30µs ? ?/sec smol-songs.csv: big filter/john 1.00 867.2±6.66µs ? ?/sec 1.01 873.3±3.44µs ? ?/sec smol-songs.csv: big filter/marcus miller 1.00 717.7±2.86µs ? ?/sec 1.01 721.5±3.89µs ? ?/sec smol-songs.csv: big filter/michael jackson 1.00 1668.4±16.76µs ? ?/sec 1.00 1667.9±10.11µs ? ?/sec smol-songs.csv: big filter/tamo 1.01 136.7±0.88µs ? ?/sec 1.00 135.5±1.22µs ? ?/sec smol-songs.csv: big filter/thelonious monk 1.03 3.1±0.02ms ? ?/sec 1.00 3.0±0.01ms ? ?/sec smol-songs.csv: desc + default/Notstandskomitee 1.00 3.0±0.01ms ? ?/sec 1.00 3.0±0.01ms ? ?/sec smol-songs.csv: desc + default/charles 1.00 1599.5±13.07µs ? ?/sec 1.01 1622.9±22.43µs ? ?/sec smol-songs.csv: desc + default/charles mingus 1.00 2.3±0.01ms ? ?/sec 1.01 2.4±0.03ms ? ?/sec smol-songs.csv: desc + default/david 1.00 5.7±0.02ms ? ?/sec 1.00 5.7±0.02ms ? ?/sec smol-songs.csv: desc + default/david bowie 1.00 9.0±0.04ms ? ?/sec 1.00 9.0±0.03ms ? ?/sec smol-songs.csv: desc + default/john 1.00 4.5±0.01ms ? ?/sec 1.00 4.5±0.02ms ? ?/sec smol-songs.csv: desc + default/marcus miller 1.00 3.9±0.01ms ? ?/sec 1.00 3.9±0.02ms ? ?/sec smol-songs.csv: desc + default/michael jackson 1.00 6.6±0.03ms ? ?/sec 1.00 6.6±0.03ms ? ?/sec smol-songs.csv: desc + default/tamo 1.00 1472.4±10.38µs ? ?/sec 1.01 1484.2±8.07µs ? ?/sec smol-songs.csv: desc + default/thelonious monk 1.00 4.4±0.02ms ? ?/sec 1.00 4.4±0.05ms ? ?/sec smol-songs.csv: desc/Notstandskomitee 1.01 2.6±0.01ms ? ?/sec 1.00 2.6±0.01ms ? ?/sec smol-songs.csv: desc/charles 1.00 475.9±3.38µs ? ?/sec 1.00 475.9±2.64µs ? ?/sec smol-songs.csv: desc/charles mingus 1.00 775.3±4.30µs ? ?/sec 1.00 778.9±3.52µs ? ?/sec smol-songs.csv: desc/david 1.00 757.9±4.10µs ? ?/sec 1.01 763.4±3.27µs ? ?/sec smol-songs.csv: desc/david bowie 1.00 1129.0±11.87µs ? ?/sec 1.01 1135.1±8.86µs ? ?/sec smol-songs.csv: desc/john 1.00 670.2±4.38µs ? ?/sec 1.00 670.2±3.46µs ? ?/sec smol-songs.csv: desc/marcus miller 1.00 961.2±4.47µs ? ?/sec 1.00 961.9±4.03µs ? ?/sec smol-songs.csv: desc/michael jackson 1.00 1076.5±6.61µs ? ?/sec 1.00 1077.9±7.11µs ? ?/sec smol-songs.csv: desc/tamo 1.00 70.6±0.57µs ? ?/sec 1.01 71.3±0.48µs ? ?/sec smol-songs.csv: desc/thelonious monk 1.01 2.9±0.01ms ? ?/sec 1.00 2.9±0.01ms ? ?/sec smol-songs.csv: prefix search/a 1.00 1236.2±9.43µs ? ?/sec 1.00 1232.0±12.07µs ? ?/sec smol-songs.csv: prefix search/b 1.00 1090.8±9.89µs ? ?/sec 1.00 1090.8±9.43µs ? ?/sec smol-songs.csv: prefix search/i 1.00 1333.9±8.28µs ? ?/sec 1.00 1334.2±11.21µs ? ?/sec smol-songs.csv: prefix search/s 1.00 810.5±3.69µs ? ?/sec 1.00 806.6±3.50µs ? ?/sec smol-songs.csv: prefix search/x 1.00 290.5±1.88µs ? ?/sec 1.00 291.0±1.85µs ? ?/sec smol-songs.csv: proximity/7000 Danses Un Jour Dans Notre Vie 1.00 4.7±0.02ms ? ?/sec 1.00 4.7±0.02ms ? ?/sec smol-songs.csv: proximity/The Disneyland Sing-Along Chorus 1.01 5.6±0.02ms ? ?/sec 1.00 5.6±0.03ms ? ?/sec smol-songs.csv: proximity/Under Great Northern Lights 1.00 2.5±0.01ms ? ?/sec 1.00 2.5±0.01ms ? ?/sec smol-songs.csv: proximity/black saint sinner lady 1.00 4.8±0.02ms ? ?/sec 1.00 4.8±0.02ms ? ?/sec smol-songs.csv: proximity/les dangeureuses 1960 1.00 3.2±0.01ms ? ?/sec 1.01 3.2±0.01ms ? ?/sec smol-songs.csv: typo/Arethla Franklin 1.00 388.7±5.16µs ? ?/sec 1.00 390.0±2.11µs ? ?/sec smol-songs.csv: typo/Disnaylande 1.01 2.6±0.01ms ? ?/sec 1.00 2.6±0.01ms ? ?/sec smol-songs.csv: typo/dire straights 1.00 125.9±1.22µs ? ?/sec 1.00 126.0±0.71µs ? ?/sec smol-songs.csv: typo/fear of the duck 1.00 373.7±4.25µs ? ?/sec 1.01 375.7±14.17µs ? ?/sec smol-songs.csv: typo/indochie 1.00 103.6±0.94µs ? ?/sec 1.00 103.4±0.74µs ? ?/sec smol-songs.csv: typo/indochien 1.00 155.6±1.14µs ? ?/sec 1.01 157.5±1.75µs ? ?/sec smol-songs.csv: typo/klub des loopers 1.00 160.6±2.98µs ? ?/sec 1.01 161.7±1.96µs ? ?/sec smol-songs.csv: typo/michel depech 1.00 79.4±0.54µs ? ?/sec 1.01 79.9±0.60µs ? ?/sec smol-songs.csv: typo/mongus 1.00 126.7±1.85µs ? ?/sec 1.00 126.1±0.74µs ? ?/sec smol-songs.csv: typo/stromal 1.01 132.9±0.99µs ? ?/sec 1.00 131.9±1.09µs ? ?/sec smol-songs.csv: typo/the white striper 1.00 287.8±2.88µs ? ?/sec 1.00 286.5±1.91µs ? ?/sec smol-songs.csv: typo/thelonius monk 1.00 304.2±1.49µs ? ?/sec 1.01 306.5±1.50µs ? ?/sec smol-songs.csv: words/7000 Danses / Le Baiser / je me trompe de mots 1.01 20.9±0.08ms ? ?/sec 1.00 20.7±0.07ms ? ?/sec smol-songs.csv: words/Bring Your Daughter To The Slaughter but now this is not part of the title 1.00 48.9±0.13ms ? ?/sec 1.00 48.9±0.11ms ? ?/sec smol-songs.csv: words/The Disneyland Children's Sing-Alone song 1.01 13.9±0.06ms ? ?/sec 1.00 13.8±0.07ms ? ?/sec smol-songs.csv: words/les liaisons dangeureuses 1793 1.01 3.7±0.01ms ? ?/sec 1.00 3.6±0.02ms ? ?/sec smol-songs.csv: words/seven nation mummy 1.00 1054.2±14.49µs ? ?/sec 1.00 1056.6±10.53µs ? ?/sec smol-songs.csv: words/the black saint and the sinner lady and the good doggo 1.00 58.2±0.29ms ? ?/sec 1.00 57.9±0.21ms ? ?/sec smol-songs.csv: words/whathavenotnsuchforth and a good amount of words to pop to match the first one 1.00 66.1±0.21ms ? ?/sec 1.00 66.0±0.24ms ? ?/sec </code> </details> Co-authored-by: Loïc Lecrenier <loic@meilisearch.com> Co-authored-by: Loïc Lecrenier <loic.lecrenier@me.com>	2022-10-25 10:42:04 +00:00
bors[bot]	cd3748d412	Merge #2951 2951: Update mini-dashboard to v0.2.3 r=curquiza a=mdubus # Pull Request ## What does this PR do? - Update the mini-dashboard with its latest release (v0.2.3) ## PR checklist Please check if your PR fulfills the following requirements: - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Morgane Dubus <30866152+mdubus@users.noreply.github.com>	2022-10-24 14:30:26 +00:00
Loïc Lecrenier	9a569d73d1	Minor code style change	2022-10-24 15:30:43 +02:00
Loïc Lecrenier	be302fd250	Remove outdated workaround for duplicate words in phrase search	2022-10-24 15:27:06 +02:00
Loïc Lecrenier	d76d0cb1bf	Merge branch 'main' into word-pair-proximity-docids-refactor	2022-10-24 15:23:00 +02:00
Morgane Dubus	079cfc70ae	Update mini-dashboard to v0.2.3	2022-10-24 15:20:59 +02:00
ManyTheFish	4afed4de4f	stabilize milli	2022-10-24 14:16:41 +02:00
ManyTheFish	a2314cf436	Update analytics	2022-10-24 13:56:26 +02:00
bors[bot]	1d85eeecef	Merge #2930 2930: fix wrong variant returned for invalid_api_key_indexes error r=ManyTheFish a=vishalsodani # Pull Request ## Related issue Fixes #2924 ## What does this PR do? - ... ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: vishalsodani <vishalsodani@rediffmail.com>	2022-10-24 11:43:03 +00:00
bors[bot]	2bf867982a	Merge #667 667: Update version for the next release (v0.34.0) in Cargo.toml files r=curquiza a=meili-bot ⚠️ This PR is automatically generated. Check the new version is the expected one before merging. Co-authored-by: curquiza <curquiza@users.noreply.github.com>	2022-10-24 10:19:04 +00:00
curquiza	f3874d58b9	Update version for the next release (v0.34.0) in Cargo.toml files	2022-10-24 10:13:25 +00:00
ManyTheFish	0578aff8c9	Fix the tests	2022-10-20 17:41:13 +02:00
ManyTheFish	1d217cef19	Add some tests	2022-10-20 17:03:07 +02:00
Loïc Lecrenier	a983129613	Apply suggestions from code review	2022-10-20 09:49:37 +02:00
bors[bot]	f11a4087da	Merge #665 665: Fixing piles of clippy errors. r=ManyTheFish a=ehiggs ## Related issue No issue fixed. Simply cleaning up some code for clippy on the march towards a clean build when #659 is merged. ## What does this PR do? Most of these are calling clone when the struct supports Copy. Many are using & and &mut on `self` when the function they are called from already has an immutable or mutable borrow so this isn't needed. I tried to stay away from actual changes or places where I'd have to name fresh variables. ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Co-authored-by: Ewan Higgs <ewan.higgs@gmail.com>	2022-10-20 07:19:46 +00:00
ManyTheFish	c02ae4dfc0	Update roaring	2022-10-19 14:25:43 +02:00
ManyTheFish	506d08a9f4	Update analytics version	2022-10-19 14:05:42 +02:00
ManyTheFish	b423ef72be	PROTO: hardcode version and interval for prototype analytics	2022-10-19 14:05:42 +02:00
ManyTheFish	77e718214f	Fix pagination analytics	2022-10-19 14:05:42 +02:00
ManyTheFish	e35ea2ad55	Make search returns 0 hits when pages is set to 0	2022-10-19 14:05:42 +02:00
ManyTheFish	dfa70e47f7	Change page and hitsPerPage corner cases	2022-10-19 14:05:42 +02:00
ManyTheFish	815fba9cc3	Fix zero division when computing pages	2022-10-19 14:05:42 +02:00
ManyTheFish	e9d493c052	Add a totalHits field on finite pagination return	2022-10-19 14:05:42 +02:00
ManyTheFish	0fa5c9b515	Fix tests	2022-10-19 14:05:42 +02:00
ManyTheFish	062d17fbc0	Use a milli version that compute exhaustivelly the number of hits	2022-10-19 14:05:42 +02:00
ManyTheFish	30410e870f	Format all fields in camelCase	2022-10-19 13:58:03 +02:00
ManyTheFish	b1bf6722e8	Update API to fit the proto needs	2022-10-19 13:58:03 +02:00
bors[bot]	f2279f4615	Merge #2928 2928: Improve default config file r=curquiza a=curquiza Following https://github.com/meilisearch/meilisearch/pull/2745#pullrequestreview-1115361274 `@gmourier` `@maryamsulemani97,` here are the changes I applied - [After discussing with the cloud team](https://meilisearch.slack.com/archives/C03T1T47TUG/p1666014688582379) (internal link only), I let commented and uncommented the variables depending on their role. - The only boolean variable I let be commented is the `no_analytics` one, and I wrote the non-default value (i.e. `true`)-> it will be easier for the users to disable analytics by only uncommenting the line. Plus, `no_analytics = false` is a double negative and not really easy to understand. - I removed the `INDEX` section that was not really related to indexes - I keep only 4 sections: SSL, Dumps, Snapshots, and an untitled section with all the uncategorized variables. I gathered all the uncategorized variable at the top of the file, so in the "untitled" section. - I copied/pasted the really basic [documentation description](https://docs.meilisearch.com/learn/configuration/instance_options.html) of the variables - I added every doc linked to each variable. Let me know if you agree or not with my choices! Thanks for your review 🙏 Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-10-19 11:49:33 +00:00
vishalsodani	1a61209596	fix wrong variant returned for invalid_api_key_indexes error	2022-10-18 19:41:06 +05:30
Clémentine Urquizar	48058b5e56	Improve default config file	2022-10-18 15:28:32 +02:00
Loïc Lecrenier	176ffd23f5	Fix compile error after rebasing wppd-refactor	2022-10-18 10:40:26 +02:00
Loïc Lecrenier	ab2f6f3aa4	Refine some details in word_prefix_pair_proximity indexing code	2022-10-18 10:37:34 +02:00
Loïc Lecrenier	e6e76fbefe	Improve performance of resolve_phrase at the cost of some relevancy	2022-10-18 10:37:34 +02:00
Loïc Lecrenier	178d00f93a	Cargo fmt	2022-10-18 10:37:34 +02:00
Loïc Lecrenier	830a7c0c7a	Use `resolve_phrase` function for exactness criteria as well	2022-10-18 10:37:34 +02:00
Loïc Lecrenier	18d578dfc4	Adjust some algorithms using DBs of word pair proximities	2022-10-18 10:37:34 +02:00
Loïc Lecrenier	072b576514	Fix proximity value in keys of prefix_word_pair_proximity_docids	2022-10-18 10:37:34 +02:00
Loïc Lecrenier	6c3a5d69e1	Update snapshots	2022-10-18 10:37:34 +02:00
Loïc Lecrenier	a7de4f5b85	Don't add swapped word pairs to the word_pair_proximity_docids db	2022-10-18 10:37:34 +02:00
Loïc Lecrenier	264a04922d	Add prefix_word_pair_proximity database Similar to the word_prefix_pair_proximity one but instead the keys are: (proximity, prefix, word2)	2022-10-18 10:37:34 +02:00
Loïc Lecrenier	1dbbd8694f	Rename StrStrU8Codec to U8StrStrCodec and reorder its fields	2022-10-18 10:37:34 +02:00
Loïc Lecrenier	bdeb47305e	Change encoding of word_pair_proximity DB to (proximity, word1, word2) Same for word_prefix_pair_proximity	2022-10-18 10:37:34 +02:00
vishalsodani	1cf6efa740	Add new error when using /keys without masterkey set	2022-10-18 10:48:45 +05:30
bors[bot]	19b2326f3d	Merge #586 586: Add settings to force milli to exhaustively compute the total number of hits r=Kerollmops a=ManyTheFish Add a new setting `exhaustive_number_hits` to `Search` forcing the `Initial` criterion to exhaustively compute the bucket_candidates allowing the end users to implement finite pagination. related to https://github.com/meilisearch/meilisearch/pull/2601 Co-authored-by: ManyTheFish <many@meilisearch.com> Co-authored-by: Many the fish <many@meilisearch.com>	2022-10-17 16:24:35 +00:00
Many the fish	81919a35a2	Update milli/src/search/criteria/initial.rs Co-authored-by: Clément Renault <clement@meilisearch.com>	2022-10-17 18:23:20 +02:00
Many the fish	516e838eb4	Update milli/src/search/criteria/initial.rs Co-authored-by: Clément Renault <clement@meilisearch.com>	2022-10-17 18:23:15 +02:00
Clément Renault	fc03e53615	Add a test to check that we can abort an indexation	2022-10-17 17:28:03 +02:00
Kerollmops	6603437cb1	Introduce an indexation abortion function when indexing documents	2022-10-17 17:28:03 +02:00
bors[bot]	96acbf815d	Merge #2913 2913: download-latest: some refactoring r=curquiza a=nfsec # Pull Request ## What does this PR do? - Usually the elevation of variables. ## PR checklist Please check if your PR fulfills the following requirements: - [X] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [X] Have you read the contributing guidelines? - [X] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Patryk Krawaczyński <nfsec@users.noreply.github.com> Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com>	2022-10-17 13:21:36 +00:00
Clémentine Urquizar - curqui	c3ecdb3d8b	Update download-latest.sh	2022-10-17 15:20:00 +02:00
ManyTheFish	6f55e7844c	Add some code comments	2022-10-17 14:41:57 +02:00
ManyTheFish	cf203b7fde	Take filter in account when computing the pages candidates	2022-10-17 14:13:44 +02:00
ManyTheFish	d71bc1e69f	Compute an exact count when using distinct	2022-10-17 14:13:44 +02:00
ManyTheFish	a396806343	Add settings to force milli to exhaustively compute the total number of hits	2022-10-17 14:13:44 +02:00
bors[bot]	fad0de4581	Merge #655 655: Upgrade all dependencies r=Kerollmops a=loiclec Upgrade all dependencies to their latest versions. Partly fixes https://github.com/meilisearch/meilisearch/issues/2822 Co-authored-by: Loïc Lecrenier <loic@meilisearch.com>	2022-10-17 11:19:46 +00:00
Loïc Lecrenier	c2ca259f48	Update cli to latest `indicatif` crate version	2022-10-17 13:05:56 +02:00
Loïc Lecrenier	4c481a8947	Upgrade all dependencies	2022-10-17 13:05:56 +02:00
bors[bot]	b0749407f3	Merge #2804 2804: Add environement variable `MEILI_CONFIG_FILE_PATH` to define the config file path r=Kerollmops a=choznerol # Pull Request ## What does this PR do? Fixes #2800 ~This is a draft PR base on the code in #2745. I will `rebase` and mark it ready for review only after #2745 merge.~ Done rebase ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? ## Demo With `config.toml`, `config_copy1.toml` and `config_copy2.toml` present: > <img width="692" alt="image" src="https://user-images.githubusercontent.com/12410942/192566891-6b3c9d26-736f-4e23-a09b-687fca1cb50d.png"> `MEILI_CONFIG_FILE_PATH` works: > <img width="773" alt="image" src="https://user-images.githubusercontent.com/12410942/192567023-f751536e-992a-4e90-a176-cb19122248be.png"> `--config-file-path` still works: > <img width="768" alt="image" src="https://user-images.githubusercontent.com/12410942/192567318-88c80b24-7873-4cec-8d08-16fe4d228055.png"> When both present, `--config-file-path` taks precedence: > <img width="1214" alt="image" src="https://user-images.githubusercontent.com/12410942/192567477-8a7cffe1-96f0-42a9-a348-6dbec20dc1e7.png"> Co-authored-by: Lawrence Chou <choznerol@protonmail.com>	2022-10-17 10:40:05 +00:00
bors[bot]	5d895dd7da	Merge #2851 2851: Upgrade clap to 4.0 r=loiclec a=choznerol # Pull Request ## Related issue Fixes #2846 This PR is draft based on #2847 to avoid conflict. I will rebase and mark as 'Ready for review' after #2847 is merged. ## What does this PR do? 1. Upgrade clap to the latest version or 4.0 (4.0.9 as of today) by following the [migrating instruction](https://github.com/clap-rs/clap/blob/master/CHANGELOG.md#migrating) from [4.0 changelog](https://github.com/clap-rs/clap/blob/master/CHANGELOG.md#migrating) 2. Fix an `ArgGroup` typo that can only be caught after upgrading to 4.0 in 20a715e29ed17c5a76229c98fb31504ada873597 ## Notable changes ### The `--help` message The format, ordering and indentation of `--help` message was changed in 4.0. I recorded the output of `cargo run -- --help` before and after upgrade to 4.0 for reference. <details> <summary>diff</summary> Output of `diff --ignore-all-space --text --unified --new-file help-message-before.txt help-message-after.txt`: ```diff --- help-message-before.txt 2022-10-14 16:45:36.000000000 +0800 +++ help-message-after.txt 2022-10-14 16:36:53.000000000 +0800 `@@` -1,12 +1,8 `@@` -meilisearch-http 0.29.1 +Usage: meilisearch [OPTIONS] -USAGE: - meilisearch [OPTIONS] - -OPTIONS: +Options: --config-file-path <CONFIG_FILE_PATH> - Set the path to a configuration file that should be used to setup the engine. Format - must be TOML + Set the path to a configuration file that should be used to setup the engine. Format must be TOML --db-path <DB_PATH> Designates the location where database files will be created and retrieved `@@` -26,15 +22,14 `@@` [default: dumps/] --env <ENV> - Configures the instance's environment. Value must be either `production` or - `development` + Configures the instance's environment. Value must be either `production` or `development` [env: MEILI_ENV=] [default: development] [possible values: development, production] -h, --help - Print help information + Print help information (use `-h` for a summary) --http-addr <HTTP_ADDR> Sets the HTTP address and port Meilisearch will use `@@` -43,63 +38,53 `@@` [default: 127.0.0.1:7700] --http-payload-size-limit <HTTP_PAYLOAD_SIZE_LIMIT> - Sets the maximum size of accepted payloads. Value must be given in bytes or explicitly - stating a base unit (for instance: 107374182400, '107.7Gb', or '107374 Mb') + Sets the maximum size of accepted payloads. Value must be given in bytes or explicitly stating a base unit (for instance: 107374182400, '107.7Gb', or '107374 Mb') [env: MEILI_HTTP_PAYLOAD_SIZE_LIMIT=] [default: 100000000] --ignore-dump-if-db-exists - Prevents a Meilisearch instance with an existing database from throwing an error when - using `--import-dump`. Instead, the dump will be ignored and Meilisearch will launch - using the existing database. + Prevents a Meilisearch instance with an existing database from throwing an error when using `--import-dump`. Instead, the dump will be ignored and Meilisearch will launch using the existing database. This option will trigger an error if `--import-dump` is not defined. [env: MEILI_IGNORE_DUMP_IF_DB_EXISTS=] --ignore-missing-dump - Prevents Meilisearch from throwing an error when `--import-dump` does not point to a - valid dump file. Instead, Meilisearch will start normally without importing any dump. + Prevents Meilisearch from throwing an error when `--import-dump` does not point to a valid dump file. Instead, Meilisearch will start normally without importing any dump. This option will trigger an error if `--import-dump` is not defined. [env: MEILI_IGNORE_MISSING_DUMP=] --ignore-missing-snapshot - Prevents a Meilisearch instance from throwing an error when `--import-snapshot` does not - point to a valid snapshot file. + Prevents a Meilisearch instance from throwing an error when `--import-snapshot` does not point to a valid snapshot file. This command will throw an error if `--import-snapshot` is not defined. [env: MEILI_IGNORE_MISSING_SNAPSHOT=] --ignore-snapshot-if-db-exists - Prevents a Meilisearch instance with an existing database from throwing an error when - using `--import-snapshot`. Instead, the snapshot will be ignored and Meilisearch will - launch using the existing database. + Prevents a Meilisearch instance with an existing database from throwing an error when using `--import-snapshot`. Instead, the snapshot will be ignored and Meilisearch will launch using the existing database. This command will throw an error if `--import-snapshot` is not defined. [env: MEILI_IGNORE_SNAPSHOT_IF_DB_EXISTS=] --import-dump <IMPORT_DUMP> - Imports the dump file located at the specified path. Path must point to a `.dump` file. - If a database already exists, Meilisearch will throw an error and abort launch + Imports the dump file located at the specified path. Path must point to a `.dump` file. If a database already exists, Meilisearch will throw an error and abort launch [env: MEILI_IMPORT_DUMP=] --import-snapshot <IMPORT_SNAPSHOT> - Launches Meilisearch after importing a previously-generated snapshot at the given - filepath + Launches Meilisearch after importing a previously-generated snapshot at the given filepath [env: MEILI_IMPORT_SNAPSHOT=] --log-level <LOG_LEVEL> Defines how much detail should be present in Meilisearch's logs. - Meilisearch currently supports five log levels, listed in order of increasing verbosity: - ERROR, WARN, INFO, DEBUG, TRACE. + Meilisearch currently supports five log levels, listed in order of increasing verbosity: ERROR, WARN, INFO, DEBUG, TRACE. [env: MEILI_LOG_LEVEL=] [default: INFO] `@@` -110,31 +95,25 `@@` [env: MEILI_MASTER_KEY=] --max-index-size <MAX_INDEX_SIZE> - Sets the maximum size of the index. Value must be given in bytes or explicitly stating a - base unit (for instance: 107374182400, '107.7Gb', or '107374 Mb') + Sets the maximum size of the index. Value must be given in bytes or explicitly stating a base unit (for instance: 107374182400, '107.7Gb', or '107374 Mb') [env: MEILI_MAX_INDEX_SIZE=] [default: 107374182400] --max-indexing-memory <MAX_INDEXING_MEMORY> - Sets the maximum amount of RAM Meilisearch can use when indexing. By default, - Meilisearch uses no more than two thirds of available memory + Sets the maximum amount of RAM Meilisearch can use when indexing. By default, Meilisearch uses no more than two thirds of available memory [env: MEILI_MAX_INDEXING_MEMORY=] [default: "21.33 TiB"] --max-indexing-threads <MAX_INDEXING_THREADS> - Sets the maximum number of threads Meilisearch can use during indexation. By default, - the indexer avoids using more than half of a machine's total processing units. This - ensures Meilisearch is always ready to perform searches, even while you are updating an - index + Sets the maximum number of threads Meilisearch can use during indexation. By default, the indexer avoids using more than half of a machine's total processing units. This ensures Meilisearch is always ready to perform searches, even while you are updating an index [env: MEILI_MAX_INDEXING_THREADS=] [default: 5] --max-task-db-size <MAX_TASK_DB_SIZE> - Sets the maximum size of the task database. Value must be given in bytes or explicitly - stating a base unit (for instance: 107374182400, '107.7Gb', or '107374 Mb') + Sets the maximum size of the task database. Value must be given in bytes or explicitly stating a base unit (for instance: 107374182400, '107.7Gb', or '107374 Mb') [env: MEILI_MAX_TASK_DB_SIZE=] [default: 107374182400] ``` - ~[help-message-before.txt](https://github.com/meilisearch/meilisearch/files/9715683/help-message-before.txt)~ [help-message-before.txt](https://github.com/meilisearch/meilisearch/files/9784156/help-message-before-2.txt) - ~[help-message-after.txt](https://github.com/meilisearch/meilisearch/files/9715682/help-message-after.txt)~ [help-message-after.txt](https://github.com/meilisearch/meilisearch/files/9784091/help-message-after.txt) </details> ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Lawrence Chou <choznerol@protonmail.com>	2022-10-17 10:08:27 +00:00
Patryk Krawaczyński	136a87f7bb	Little code refactor Usually the elevation of variables.	2022-10-14 15:32:59 +02:00
Lawrence Chou	5dafdd9a23	Preserve --help output ordering after upgrade Clap to 4.0 From the [4.0 breaking change][1]: ... * (help) Make DeriveDisplayOrder the default and removed the setting. To sort help, set next_display_order(None) (#2808) ... [1]: https://github.com/clap-rs/clap/blob/master/CHANGELOG.md#breaking-changes	2022-10-14 16:43:03 +08:00
Lawrence Chou	53e5229b4a	Assert error message for Windows besides *nix The 'Tests on windows-latest' now failed with error message below ---- option::test::test_meilli_config_file_path_invalid stdout ---- thread 'option::test::test_meilli_config_file_path_invalid' panicked at 'assertion failed: left: `"unable to open or read the \"../configgg.toml\" configuration file: The system cannot find the file specified. (os error 2)."`, right: `"unable to open or read the \"../configgg.toml\" configuration file: No such file or directory (os error 2)."`', meilisearch-http\src\option.rs:555:17 https://github.com/meilisearch/meilisearch/actions/runs/3231941308/jobs/5291998750	2022-10-14 14:49:40 +08:00
Lawrence Chou	9ebc73e6ac	Comply with Clippy rule single-match	2022-10-14 14:16:10 +08:00
Ewan Higgs	beb987d3d1	Fixing piles of clippy errors. Most of these are calling clone when the struct supports Copy. Many are using & and &mut on `self` when the function they are called from already has an immutable or mutable borrow so this isn't needed. I tried to stay away from actual changes or places where I'd have to name fresh variables.	2022-10-13 22:02:54 +02:00
bors[bot]	fc5a7e376c	Merge #2876 2876: Full support for compressed (Gzip, Brotli, Zlib) API requests r=Kerollmops a=mou # Pull Request ## Related issue Fixes #2802 ## What does this PR do? - Adds missed content-encoding support for streamed requests (documents) - Adds additional tests to validate content-encoding support is in place - Adds new tests to validate content-encoding support for previously existing code (built-in actix functionality for unmarshaling JSON payloads) ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Andrey "MOU" Larionov <anlarionov@gmail.com>	2022-10-13 15:32:21 +00:00
bors[bot]	5a98ecaf2a	Merge #2896 2896: Fix CI to send signal to Cloud team r=Kerollmops a=curquiza Hello `@eskombro` 👋 I realized the CI we recently created together was not good when we release the official Meilisearch version. Indeed, in this case `steps.meta.outputs.tags` contains several tags, see: https://github.com/meilisearch/meilisearch/actions/runs/3197492898/jobs/5220776456 You might want to ask: why the CI, including the Cloud team signal, was available when releasing v0.29.1 and not v0.29.0? Good question, thanks for asking it Sam! It was a mistake on our side, it should not have been available for v0.29.1, and this is how I found out v0.29.1 was broken and contained commits it should not have. So I deleted everything and started the release process again for v0.29.1. Anyway, since I had the chance to see the bug in this release mess, I want to take the opportunity to fix it. Now, we will only send the real tag. Here you have more documentation about what `github.ref_name` is: https://docs.github.com/en/actions/learn-github-actions/contexts Bisous bisous! Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com>	2022-10-13 14:17:07 +00:00
Andrey "MOU" Larionov	2dce44f4c1	Fix formatting and shaving lints	2022-10-13 15:49:23 +02:00
bors[bot]	95e45e1c2c	Merge #663 663: Fix CONTRIBUTING.md step to make the project work r=Kerollmops a=curquiza Following this discussion: https://github.com/meilisearch/milli/issues/76#issuecomment-1277459125 Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com>	2022-10-13 11:47:34 +00:00
Clémentine Urquizar - curqui	59fe1e8efa	Update CONTRIBUTING.md	2022-10-13 13:46:18 +02:00
bors[bot]	f30979d021	Merge #662 662: Enhance word splitting strategy r=ManyTheFish a=akki1306 # Pull Request ## Related issue Fixes #648 ## What does this PR do? - [split_best_frequency](`55d889522b/milli/src/search/query_tree.rs (L282-L301)`) to use frequency of word pairs near together with proximity value of 1 instead of considering the frequency of individual words. Word pairs having max frequency are considered. ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Akshay Kulkarni <akshayk.gj@gmail.com>	2022-10-13 08:14:22 +00:00
Akshay Kulkarni	85f3028317	remove underscore and introduce back word_documents_count	2022-10-13 13:21:59 +05:30
Akshay Kulkarni	8195fc6141	revert removal of word_documents_count method	2022-10-13 13:14:27 +05:30
Akshay Kulkarni	32f825d442	move default implementation of word_pair_frequency to TestContext	2022-10-13 12:57:50 +05:30
Akshay Kulkarni	ff8b2d4422	formatting	2022-10-13 12:44:08 +05:30
Akshay Kulkarni	6cb8b46900	use word_pair_frequency and remove word_documents_count	2022-10-13 12:43:11 +05:30
Andrey "MOU" Larionov	b69f8d67c3	Added test to verify response encoding Alongside request encoding (compression) support, it is helpful to verify that the server respect `Accept-Encoding` headers and apply the corresponding compression to responses.	2022-10-13 00:56:57 +02:00
Andrey "MOU" Larionov	99e2788ee7	Fix Cargo.toml formatting	2022-10-12 21:12:18 +02:00
Akshay Kulkarni	8c9245149e	format file	2022-10-12 15:27:56 +05:30
bors[bot]	2000f7958d	Merge #604 604: Speed up debug builds r=Kerollmops a=loiclec Note: this draft PR is based on https://github.com/meilisearch/milli/pull/601 , for no particular reason. ## What does this PR do? Make a series of changes with the goal of speeding up debug builds: 1. Add an `all_languages` feature which compiles charabia with its `default` features activated. The `all_languages` feature is activated by default. But running: ``` cargo build --no-default-features ``` on `milli` is now much faster. 2. Reduce the debug optimisation level from 3 to 0, except for a few critical dependencies. 3. Compile the build dependencies quicker as well. Previously, all build dependencies were compiled with `opt-level = 3`. Now, only the critical build dependencies are compiled with optimisations. 4. Reduce the amount of code generated by the `documents!` macro 5. Make the "progress update" closure provided to indexing functions a trait object instead of a generic parameter. This avoids monomorphising the indexing code multiple times needlessly. ## Results Initial build times on my computer before and after these changes: \| \| cargo check \| cargo check --no-default-features \| cargo test \| cargo test --lib \| cargo test --no-default-features \| cargo test --lib --no-default-features \| \|--------\|-------------\|-----------------------------------\|------------\|------------------\|----------------------------------\|----------------------------------------\| \| before \| 1m05s \| 1m05s \| 2m06s \| 1m47s \| 2m06 \| 1m47s \| \| after \| 28.9s \| 13.1s \| 40s \| 38s \| 23s \| 21s \| Co-authored-by: Loïc Lecrenier <loic@meilisearch.com>	2022-10-12 08:54:48 +00:00
Akshay Kulkarni	63e79a9039	update comment	2022-10-12 13:36:48 +05:30
Akshay Kulkarni	7f9680f0a0	Enhance word splitting strategy	2022-10-12 13:18:23 +05:30
Clémentine Urquizar - curqui	a9e6a8901b	Fix CI to send signal to Cloud team	2022-10-12 09:36:01 +02:00
Loïc Lecrenier	53503f09ca	Make milli's default features optional in other executable targets	2022-10-12 09:22:05 +02:00
Loïc Lecrenier	6fbf5dac68	Simplify documents! macro to reduce compile times	2022-10-12 09:22:05 +02:00
Loïc Lecrenier	98fc093823	Optimize a few performance sensitive dependencies on debug builds	2022-10-12 09:22:05 +02:00
Loïc Lecrenier	5cfb5df31e	Set opt-level to 0 for debug builds But speed up compile times by optimising build dependencies of lindera	2022-10-12 09:22:05 +02:00
Lawrence Chou	3c3ae3ff98	Impeove invalid config_file_path handling 1. Besides opt.config_file_path, also consider MEILI_CONFIG_FILE_PATH in the Err path because they are both user input. 2. Print out the incorrect file path in error message. 3. Add tests https://github.com/meilisearch/meilisearch/pull/2804#discussion_r991999888	2022-10-12 12:04:48 +08:00
bors[bot]	343828a76e	Merge #2890 2890: fix: add handle dumpCreation query on tasks request r=Kerollmops a=washbin # Pull Request ## Related issue Fixes #2874 ## What does this PR do? - add missing `DumpCreation` type in tasks route ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Co-authored-by: washbin <76929116+washbin@users.noreply.github.com>	2022-10-11 19:09:34 +00:00
washbin	72c1aef1c4	fix: add handle dumpCreation query on tasks request	2022-10-11 19:36:04 +05:45
Lawrence Chou	91accc0194	Fix default config file path typo	2022-10-11 21:36:17 +08:00
bors[bot]	0f024e7d97	Merge #2882 2882: Bring back v0.29.1 changes into `main` r=Kerollmops a=curquiza Following v0.29.1 release Co-authored-by: Loïc Lecrenier <loic@meilisearch.com> Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com> Co-authored-by: bors[bot] <26634292+bors[bot]@users.noreply.github.com>	2022-10-10 16:27:22 +00:00
bors[bot]	55d889522b	Merge #658 658: Add proximity calculation for the same word r=ManyTheFish a=msvaljek # Pull Request ## Related issue Fixes https://github.com/meilisearch/milli/issues/647 ## What does this PR do? - During [the increase of the current word position](`d94339a858/milli/src/update/index_documents/extract/extract_word_pair_proximity_docids.rs (L129-L135)`) we extract the proximity between the current position and the next one. ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: msvaljek <marko.svaljek@commercetools.com>	2022-10-10 13:33:58 +00:00
Clémentine Urquizar	3ebd88c03b	Revert "Comment cache steps in jobs" This reverts commit `f513ac1233`.	2022-10-10 14:46:54 +02:00
bors[bot]	c958097e99	Merge #2862 2862: Use Ubuntu 18.04 for all CI tasks that previously used Ubuntu 20.04 r=curquiza a=loiclec This is to prevent linking with a version of glibc that is too recent. With meilisearch v0.29.0 we inadvertently bumped the minimum supported glibc version to 2.29, which means it couldn't be run from Debian 10 (for example) anymore. By using Ubuntu 18.04, which uses glibc 2.27, we restore support for older Linux distros. Fixes #2850 Co-authored-by: Loïc Lecrenier <loic@meilisearch.com> Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-10-10 14:42:18 +02:00
Clémentine Urquizar	f513ac1233	Comment cache steps in jobs	2022-10-10 14:25:24 +02:00
Loïc Lecrenier	97c202db51	Update version for next release (v0.29.1)	2022-10-10 14:25:24 +02:00
Loïc Lecrenier	a5e23aa6e4	Use Ubuntu 18.04 for all CI tasks that previously used Ubuntu 20.04 This is to prevent linking with a version of glibc that is too recent. With meilisearch v0.29.0 we inadvertently bumped the minimum supported glibc version to 2.29, which means it couldn't be run from Debian 10 (for example) anymore. By using Ubuntu 18.04, which uses glibc 2.27, we restore support for older Linux distros.	2022-10-10 14:25:18 +02:00
bors[bot]	c5cd743eb6	Merge #2868 2868: Uncomment cache steps in Github CI r=curquiza a=AM1TB # Pull Request Uncomment cache steps as they were previously affected by an issue with Github actions: https://www.githubstatus.com/incidents/gq1x0j8bv67v ## Related issue Fixes #2864 ## What does this PR do? - Reintroduce the rust-cache steps within Github CI. ## PR checklist Please check if your PR fulfills the following requirements: - [X] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [X] Have you read the contributing guidelines? - [X] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Amit Banerjee <amit.banerjee.jsr@gmail.com>	2022-10-10 11:33:37 +00:00
Andrey "MOU" Larionov	343a677566	Fix formatting and apply clippy suggestion	2022-10-10 11:04:46 +02:00
Andrey "MOU" Larionov	9dbc71cb6d	Added support for encoded payload Actix provides different content encodings out of the box, but only if we use built-in content wrappers and containers. This patch wraps its own Payload implementation with an actix decoder, which enables request compression support.	2022-10-09 22:09:30 +02:00
Andrey "MOU" Larionov	11b986a81d	Added support for specifying compression in tests Refactored tests code to allow to specify compression (content-encoding) algorithm. Added tests to verify what actix actually handle different content encodings properly.	2022-10-09 22:09:29 +02:00
Andrey "MOU" Larionov	7607a62531	Split tests over two modules Currently, `add_documents` contains some amount of `update` tests. This change should unify test structure with `index` module.	2022-10-09 22:03:22 +02:00
bors[bot]	c883b23cca	Merge #2861 2861: Change default bind address to localhost r=Kerollmops a=Fall1ngStar # Pull Request ## Related issue Fixes #2782 ## What does this PR do? - Change the default bind address to `localhost` so that it can be accessed with IPv6 ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Fall1ngStar <fall1ngstar.public@gmail.com>	2022-10-07 19:49:39 +00:00
msvaljek	762e320c35	Add proximity calculation for the same word	2022-10-07 12:59:12 +02:00
Fall1ngStar	d1c10d6d68	Fix segment_analytics default_http_addr import	2022-10-06 22:42:20 -04:00
Amit Banerjee	1f40c3e48c	Uncomment cache steps in jobs	2022-10-06 22:32:29 +05:30
Lawrence Chou	2681e92d4e	Support MEILI_CONFIG_FILE_PATH to define config file path Close #2800 This is an alternative to the `--config-file-path` option. If both `--config-file-path` and `MEILI_CONFIG_FILE_PATH` are present, `--config-file-path` takes precedence according to the "Priority order" section of #2558.	2022-10-07 00:41:14 +08:00
Lawrence Chou	da25328c2b	Fix clap ArgGroup typos Not sure why but the compiler didn't catch this until clap is upgraded to v4. Follwoing are the error from 'cargo test': running 2 tests test routes::indexes::search::test::test_fix_sort_query_parameters ... ok test option::test::test_valid_opt ... FAILED failures: ---- option::test::test_valid_opt stdout ---- thread 'option::test::test_valid_opt' panicked at 'Command meilisearch-http: Argument or group 'import-snapshot' specified in 'requires*' for 'ignore_missing_snapshot' does not exist', /Users/ychou/.cargo/registry/src/github.com-1ecc6299db9ec823/clap-4.0.9/src/builder/debug_asserts.rs:152:13 note: run with environment variable to display a backtrace failures: option::test::test_valid_opt test result: FAILED. 1 passed; 1 failed; 0 ignored; 0 measured; 0 filtered out; finished in 0.00s	2022-10-07 00:32:26 +08:00
Lawrence Chou	9e5ef8eb69	Upgrade clap to v4 Close #2846 4.0.0 changelog: 'https://github.com/clap-rs/clap/blob/master/CHANGELOG.md#400---2022-09-28' I followed the [Migrating steps](https://github.com/clap-rs/clap/blob/master/CHANGELOG.md#migrating) and the only issue I encountered are: 1. The typo problem in previous commit "Fix clap ArgGroup typo" 2. I can't say I am 100% sure every [Subtle changes](https://github.com/clap-rs/clap/blob/master/CHANGELOG.md#breaking-changes) is fine for our use case, but at least after a quick read I didn't notice anything actionable.	2022-10-07 00:32:25 +08:00
Lawrence Chou	6285c5949c	Fix clap v4 deprecation warning Following the 3. stap of [clap v4 migration instructions](https://github.com/clap-rs/clap/blob/master/CHANGELOG.md#migrating) The result for 'cargo check --features clap/deprecated' is https://user-images.githubusercontent.com/12410942/193825216-ac680574-f53b-49c0-88c4-8bc42c4c6381.png	2022-10-07 00:15:53 +08:00
Lawrence Chou	b55ec7db4d	Upgrade clap to 3.2.8 Upgrade to the latest version of v3 before upgrading to v4	2022-10-07 00:04:21 +08:00
bors[bot]	1b72eba1f3	Merge #2867 2867: Bring back `stable` into `main` r=Kerollmops a=curquiza Following hotfix for v0.29.1 Co-authored-by: Loïc Lecrenier <loic@meilisearch.com> Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com> Co-authored-by: bors[bot] <26634292+bors[bot]@users.noreply.github.com>	2022-10-06 14:14:23 +00:00
bors[bot]	c98d3209ad	Merge #2839 2839: Update internal CLI documentation r=curquiza a=jeertmans # Pull Request ## Related issue Fixes #2810 ## What does this PR do? - Make internal CLI documentation match the online one ## Remarks - Scope is limited to `meilisearch/meilisearch-http/src/option.rs`, not the flattened structs: should I also take care of them? - Could not find online docs for `enable_metrics_route` - Max column width wrapping was done by hand, so may not be perfect - I removed the links from the internal doc: should I put them back? ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Co-authored-by: Jérome Eertmans <jeertmans@icloud.com>	2022-10-06 13:32:32 +00:00
Jérome Eertmans	abb0233077	chore: add docs of flattened structs	2022-10-06 15:14:56 +02:00
Jérome Eertmans	8c526c31da	Update meilisearch-http/src/option.rs Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com>	2022-10-06 15:08:37 +02:00
bors[bot]	ed6c2fb22c	Merge #2862 2862: Use Ubuntu 18.04 for all CI tasks that previously used Ubuntu 20.04 r=curquiza a=loiclec This is to prevent linking with a version of glibc that is too recent. With meilisearch v0.29.0 we inadvertently bumped the minimum supported glibc version to 2.29, which means it couldn't be run from Debian 10 (for example) anymore. By using Ubuntu 18.04, which uses glibc 2.27, we restore support for older Linux distros. Fixes #2850 Co-authored-by: Loïc Lecrenier <loic@meilisearch.com> Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-10-06 12:18:38 +00:00
Clémentine Urquizar	ab17c0acd5	Comment cache steps in jobs	2022-10-06 14:03:04 +02:00
bors[bot]	425692287d	Merge #2841 2841: Bail if config file contains 'config_file_path' r=Kerollmops a=arriven # Pull Request ## Related issue Fixes #2801 ## What does this PR do? - Return an error if config file contains 'config_file_path' ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Co-authored-by: arriven <20084245+Arriven@users.noreply.github.com>	2022-10-06 11:58:28 +00:00
Loïc Lecrenier	05af8f0e46	Update version for next release (v0.29.1)	2022-10-06 10:27:11 +02:00
Loïc Lecrenier	dc93853946	Use Ubuntu 18.04 for all CI tasks that previously used Ubuntu 20.04 This is to prevent linking with a version of glibc that is too recent. With meilisearch v0.29.0 we inadvertently bumped the minimum supported glibc version to 2.29, which means it couldn't be run from Debian 10 (for example) anymore. By using Ubuntu 18.04, which uses glibc 2.27, we restore support for older Linux distros.	2022-10-06 10:13:50 +02:00
Fall1ngStar	435778f328	Change default bind address to localhost	2022-10-05 22:23:29 -04:00
bors[bot]	8134c26a6f	Merge #2847 2847: Upgrade dependencies r=Kerollmops a=loiclec # Pull Request ## Related issue Partly fixes #2822 ## What does this PR do? Upgrade Meilisearch's dependencies to their latest versions, except: - clap stays on 3.0 because it is more complicated to upgrade ( see #2846 ) - `enum_iterator` goes up to 1.1.2 instead of 1.2 because of the vergen dependency ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Co-authored-by: Loïc Lecrenier <loic@meilisearch.com>	2022-10-05 20:09:17 +00:00
bors[bot]	358aa337ea	Merge #657 657: Fix link in Hacktoberfest section r=curquiza a=meili-bot _This PR is auto-generated._ Fix link in CONTRIBUTING.md. Following [this PR](https://github.com/meilisearch/meilisearch/pull/2845) and [this issue](https://github.com/meilisearch/meilisearch/issues/2840). Co-authored-by: meili-bot <74670311+meili-bot@users.noreply.github.com>	2022-10-05 17:19:33 +00:00
meili-bot	1764a33690	Update CONTRIBUTING.md	2022-10-05 19:19:03 +02:00
Jérome Eertmans	6933ba54fb	chore: add missing docs	2022-10-05 19:04:33 +02:00
Jérome Eertmans	4a022acd33	Update meilisearch-http/src/option.rs Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com> Update meilisearch-http/src/option.rs Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com> Update meilisearch-http/src/option.rs Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com> Update meilisearch-http/src/option.rs Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com> Update meilisearch-http/src/option.rs Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com> Update meilisearch-http/src/option.rs Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com> Update meilisearch-http/src/option.rs Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com> Update meilisearch-http/src/option.rs Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com> Update meilisearch-http/src/option.rs Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com> Update meilisearch-http/src/option.rs Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com> Update meilisearch-http/src/option.rs Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com> Apply suggestions from code review Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com>	2022-10-05 18:54:38 +02:00
Loïc Lecrenier	a2c91a87fe	Upgrade dependencies Except: - clap stays on 3.0 because it is more complicated to upgrade - enum_iterator goes up to 1.1.2 instead of 1.2 because of the vergen dependency	2022-10-05 15:53:02 +02:00
arriven	59f1091c5e	Bail if config file contains 'config_file_path'	2022-10-05 12:00:27 +03:00
Jérome Eertmans	d1fdca2bce	chore(fmt): cargo fmt	2022-10-05 10:31:29 +02:00
bors[bot]	9da34dc55c	Merge #2837 2837: chore: generate Apple Silicon binaries r=curquiza a=jeertmans # Pull Request This creates a new job in the "publish binaries" workflow. User `@mohitsaxenaknoldus` did not seem to be active on this anymore, so I tryied myself ;-) ## Related issue Fixes #2792 ## What does this PR do? - As titled ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Co-authored-by: Jérome Eertmans <jeertmans@icloud.com>	2022-10-05 07:42:02 +00:00
Jérome Eertmans	221e3edf48	Update publish-binaries.yml	2022-10-04 19:33:39 +02:00
Clémentine Urquizar - curqui	add793462f	Update milestone-workflow.yml (#2853 )	2022-10-04 16:35:06 +02:00
Clémentine Urquizar - curqui	cc2a271287	Update milestone-workflow.yml (#2852 )	2022-10-04 16:26:15 +02:00
bors[bot]	77cdecaf3a	Merge #2848 2848: v0.29.0: bring `stable` changes into `main` r=curquiza a=curquiza Bring `stable` changes into `main`. These changes correspond to changes we made during the v0.29.0 pre-release Co-authored-by: evpeople <hangcaihui@gmail.com> Co-authored-by: Loïc Lecrenier <loic@meilisearch.com> Co-authored-by: bors[bot] <26634292+bors[bot]@users.noreply.github.com> Co-authored-by: ManyTheFish <many@meilisearch.com> Co-authored-by: Many the fish <many@meilisearch.com> Co-authored-by: Morgane Dubus <30866152+mdubus@users.noreply.github.com> Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com> Co-authored-by: Kerollmops <clement@meilisearch.com> Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com>	2022-10-04 12:29:40 +00:00
Clémentine Urquizar - curqui	a7d2c9572e	Merge branch 'main' into stable	2022-10-04 14:20:10 +02:00
bors[bot]	a90d7e4cc7	Merge #654 654: Re-upload milli's logo r=curquiza a=jeertmans # Pull Request ## Related issue None ## What does this PR do? Apparently, some [commit](`add96f921b`) deleted the logo file, and updated the `src` path. It seems to me that this was an error, and that the logo file should have been moved, not deleted. This fixes the problem of seeing this (see image) instead of the actual logo. ![image](https://user-images.githubusercontent.com/27275099/193786803-e0d11a59-48fa-4331-bd92-48457969d766.png) ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Co-authored-by: Jérome Eertmans <jeertmans@icloud.com>	2022-10-04 10:56:33 +00:00
Jérome Eertmans	aec220ab63	chore: move logo to (new) assets folder	2022-10-04 12:20:24 +02:00
bors[bot]	53c7c1b07f	Merge #2845 2845: Fixed broken Link r=curquiza a=AnirudhDaya # Pull Request ## Related issue Fixes #2840 ## What does this PR do? - Broken link is fixed. Now it redirects to the correct website. ## PR checklist Please check if your PR fulfils the following requirements: - [X] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [X] Have you read the contributing guidelines? - [X] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Anirudh Dayanand <anirudhdaya@gmail.com>	2022-10-04 10:04:31 +00:00
Jérome Eertmans	4348c49656	fix: re-upload milli's logo The logo was deleted with this [commit](`add96f921b`).	2022-10-04 11:33:19 +02:00
Jérome Eertmans	a93e3dead3	Update meilisearch-http/src/option.rs Co-authored-by: Tommy <68053732+dichotommy@users.noreply.github.com>	2022-10-04 10:30:08 +02:00
Anirudh Dayanand	62b06e6088	Fixed broken Link	2022-10-04 13:39:40 +05:30
bors[bot]	a18de9b5f0	Merge #650 650: Add missing logging timer to extractors r=Kerollmops a=vishalsodani # Pull Request ## What does this PR do? #645 <!-- Please link the issue you're trying to fix with this PR, if none then please create an issue first. --> ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: vishalsodani <vishalsodani@rediffmail.com>	2022-10-04 07:25:47 +00:00
bors[bot]	8ad5c3043c	Merge #2819 2819: Replace a meaningless serde message r=irevoire a=onyxcherry # Pull Request ## What does this PR do? Fixes #2680 <!-- Please link the issue you're trying to fix with this PR, if none then please create an issue first. --> I've renamed the `serde_msg` variable to a `message` as _message_ does or does not include the serde error message --> is more generic. ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Tomasz Wiśniewski <tomasz@wisniewski.app>	2022-10-03 17:58:00 +00:00
bors[bot]	8fa8eb9767	Merge #2830 2830: Increase max concurrent readers on indexes r=irevoire a=arriven # Pull Request ## What does this PR do? Fixes #2786 <!-- Please link the issue you're trying to fix with this PR, if none then please create an issue first. --> ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: arriven <20084245+Arriven@users.noreply.github.com>	2022-10-03 17:26:45 +00:00
bors[bot]	069374c23a	Merge #2817 2817: Update Hacktoberfest section in CONTRIBUTING.md r=curquiza a=meili-bot _This PR is auto-generated._ Following: `af850854e4` Update Hacktoberfest section in CONTRIBUTING.md with the global guideline information. Co-authored-by: meili-bot <74670311+meili-bot@users.noreply.github.com>	2022-10-03 14:49:50 +00:00
Jérome Eertmans	458775c7ad	docs: add missing options	2022-10-03 16:37:16 +02:00
Jérome Eertmans	23db798f45	docs: update CLI documentation Closes #2810	2022-10-03 16:07:38 +02:00
Jérome Eertmans	6a67afa48b	chore: generate Apple Silicon binaries Closes #2792	2022-10-03 15:09:39 +02:00
bors[bot]	4ccfc837c1	Merge #2827 2827: Upgrade to alpine 3.16 r=curquiza a=nontw # Pull Request ## What does this PR do? Just a simple version upgrade to the existing Dockerfile. Tested building the image locally successfully. ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: nont <nontkrub@gmail.com>	2022-10-03 12:28:16 +00:00
bors[bot]	ce75bb5ef8	Merge #2834 2834: Deleted v1.rs because it is not included in Project r=irevoire a=Himanshu664 # Pull Request ## Related issue Fixes #2833 ## What does this PR do? - Deleted the v1.rs file because it is not included in project/ ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue, or have you listed the changes applied in the PR description (and why they are needed)? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Himanshu Malviya <gs2012023@sgsitsindore.in>	2022-10-03 11:10:05 +00:00
Himanshu Malviya	f7f34fb714	deleted v1.rs	2022-10-03 11:04:54 +00:00
bors[bot]	fcedce8578	Merge #2745 #2789 #2814 #2826 2745: Config file support r=curquiza a=mlemesle # Pull Request ## What does this PR do? Fixes #2558 ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! 2789: Fix typos r=Kerollmops a=kianmeng # Pull Request ## What does this PR do? Found via `codespell -L crate,nam,hart`. ## PR checklist Please check if your PR fulfills the following requirements: - [ ] Does this PR fix an existing issue? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! 2814: Skip dashboard test if mini-dashboard feature is disabled r=Kerollmops a=jirutka Fixes #2813 Fixes the following error: cargo test --no-default-features ... error: couldn't read target/debug/build/meilisearch-http-ec029d8c902cf2cb/out/generated.rs: No such file or directory (os error 2) --> meilisearch-http/tests/dashboard/mod.rs:8:9 \| 8 \| include!(concat!(env!("OUT_DIR"), "/generated.rs")); \| ^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^ \| = note: this error originates in the macro `include` (in Nightly builds, run with -Z macro-backtrace for more info) error: could not compile `meilisearch-http` due to previous error 2826: Rename receivedDocumentIds into matchedDocuments r=Kerollmops a=Ugzuzg # Pull Request ## What does this PR do? Fixes #2799 Changes DocumentDeletion task details response. ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Tested with curl: ``` curl \ -X POST 'http://localhost:7700/indexes/movies/documents/delete-batch' \ -H 'Content-Type: application/json' \ --data-binary '[ 23488, 153738, 437035, 363869 ]' {"taskUid":1,"indexUid":"movies","status":"enqueued","type":"documentDeletion","enqueuedAt":"2022-10-01T20:06:37.105416054Z"}% curl \ -X GET 'http://localhost:7700/tasks/1' {"uid":1,"indexUid":"movies","status":"succeeded","type":"documentDeletion","details":{"matchedDocuments":4,"deletedDocuments":2},"duration":"PT0.005708322S","enqueuedAt":"2022-10-01T20:06:37.105416054Z","startedAt":"2022-10-01T20:06:37.115562733Z","finishedAt":"2022-10-01T20:06:37.121271055Z"} ``` Co-authored-by: mlemesle <lemesle.martin@hotmail.fr> Co-authored-by: Kian-Meng Ang <kianmeng@cpan.org> Co-authored-by: Jakub Jirutka <jakub@jirutka.cz> Co-authored-by: Jarasłaŭ Viktorčyk <ugzuzg@gmail.com>	2022-10-03 09:48:34 +00:00
bors[bot]	a93eec1d34	Merge #2831 2831: Make clippy happy r=curquiza a=Kerollmops Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-10-03 08:48:07 +00:00
Kerollmops	135f656e8f	Make clippy happy	2022-10-03 10:39:42 +02:00
bors[bot]	f9c2dacf33	Merge #653 653: Fix #652 - Change Spelling of `author` in `README.md` r=curquiza a=anirudhRowjee # Pull Request ## What does this PR do? Fixes #652 - Changes spellings of `au{hor` to `author` - Minor formatting changes in Markdown ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Co-authored-by: Anirudh Rowjee <ani.rowjee@gmail.com>	2022-10-03 08:20:48 +00:00
Anirudh Rowjee	7d247353d0	[docs] contd - fix #652 , revert capitalization of 'Meilisearch'	2022-10-03 09:52:20 +05:30
Anirudh Rowjee	bc502ee125	[docs] Fixed #652 , changes spelling of author	2022-10-03 09:38:59 +05:30
arriven	88e69f4302	Increase max concurrent readers on indexes	2022-10-02 18:25:32 +03:00
nont	459829631f	Upgrade to alpine 3.16	2022-10-01 18:06:09 -07:00
Jarasłaŭ Viktorčyk	20589a41b5	Rename receivedDocumentIds into matchedDocuments Changes DocumentDeletion task details response.	2022-10-01 21:59:20 +02:00
vishalsodani	00c02d00f3	Add missing logging timer to extractors	2022-09-30 22:17:06 +05:30
bors[bot]	804db03e41	Merge #649 649: Update Hacktoberfest section in CONTRIBUTING.md r=curquiza a=meili-bot _This PR is auto-generated._ Following: `af850854e4` Update Hacktoberfest section in CONTRIBUTING.md with the global guideline information. Co-authored-by: meili-bot <74670311+meili-bot@users.noreply.github.com>	2022-09-29 15:50:20 +00:00
Tomasz Wiśniewski	61a518a384	Fix #2680 - replace a meaningless serde message	2022-09-29 16:36:32 +02:00
meili-bot	26efdf4dd9	Update CONTRIBUTING.md	2022-09-29 16:00:15 +02:00
meili-bot	7905dae7ad	Update CONTRIBUTING.md	2022-09-29 16:00:08 +02:00
Jakub Jirutka	05f93541d8	Skip dashboard test if mini-dashboard feature is disabled Fixes the following error: cargo test --no-default-features ... error: couldn't read target/debug/build/meilisearch-http-ec029d8c902cf2cb/out/generated.rs: No such file or directory (os error 2) --> meilisearch-http/tests/dashboard/mod.rs:8:9 \| 8 \| include!(concat!(env!("OUT_DIR"), "/generated.rs")); \| ^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^ \| = note: this error originates in the macro `include` (in Nightly builds, run with -Z macro-backtrace for more info) error: could not compile `meilisearch-http` due to previous error	2022-09-29 01:42:52 +02:00
bors[bot]	4b903719a0	Merge #643 643: Add Hacktoberfest section to CONTRIBUTING.md r=curquiza a=meili-bot _This PR is auto-generated._ Add Hacktoberfest section to CONTRIBUTING.md Co-authored-by: meili-bot <74670311+meili-bot@users.noreply.github.com>	2022-09-22 16:44:51 +00:00
meili-bot	ed3d87f061	Update CONTRIBUTING.md	2022-09-22 18:43:42 +02:00
Clémentine Urquizar - curqui	2827ff7957	Update README.md with Hacktoberfest section (#2794 ) * Update README.md * Update README.md * Update README.md Co-authored-by: Bruno Casali <brunoocasali@gmail.com> Co-authored-by: Bruno Casali <brunoocasali@gmail.com>	2022-09-22 17:45:33 +02:00
Luna-meili	d166a97d67	Update CONTRIBUTING.md for Hacktoberfest (#2793 ) * Update CONTRIBUTING.md * Update CONTRIBUTING.md Co-authored-by: Bruno Casali <brunoocasali@gmail.com> * Update CONTRIBUTING.md Co-authored-by: Bruno Casali <brunoocasali@gmail.com> Co-authored-by: Bruno Casali <brunoocasali@gmail.com>	2022-09-22 17:27:42 +02:00
mlemesle	248d727e61	Add quotes around file name and change error message	2022-09-22 09:44:28 +02:00
mlemesle	56d72d4493	Uncomment static default values and fix typo	2022-09-21 16:31:16 +02:00
Kian-Meng Ang	740926e747	Fix typos Found via `codespell -L crate,nam,hart,succeded`.	2022-09-21 21:46:06 +08:00
bors[bot]	f4b81fa0a1	Merge #2772 #2773 2772: Improve issue template display to avoid support in Meilisearch issues r=curquiza a=curquiza Move the support line higher to make it more visible Like this: https://github.com/meilisearch/meilisearch/discussions/2780 2773: Allow building without specialized tokenizations r=curquiza a=jirutka Fixes #2774 (Some of) these specialized tokenizations include huge dictionaries that currently account for 90% (!) of the meilisearch binary size. This commit adds `chinese`, `hebrew`, `japanese`, and `thai` feature flags that are propagated via `milli` down to the `charabia` crate. To keep it backwards compatible, they are enabled by default. Related to meilisearch/milli#632 Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com> Co-authored-by: Jakub Jirutka <jakub@jirutka.cz>	2022-09-21 10:56:12 +00:00
bors[bot]	0fdbf128e9	Merge #2790 2790: Improve Docker CI for cloud team r=curquiza a=curquiza Work with `@eskombro` Update CI to send a signal to Cloud team CI when Docker image is pushed Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-09-21 09:55:50 +00:00
Clémentine Urquizar	d3b984d862	Update CI to send a signal to Cloud team when Docker image is pushed Co-authored-by: Samuel Jimenez <sjimenezre@gmail.com>	2022-09-21 11:39:58 +02:00
bors[bot]	a3622eda46	Merge #642 642: Remove LTO in release profile r=Kerollmops a=loiclec Since we can't enable it in Meilisearch (see https://github.com/meilisearch/meilisearch/pull/2717 ), we should not enable it in milli either. The goal is for milli's benchmarks to accurately represent its performance within meilisearch. Co-authored-by: Loïc Lecrenier <loic@meilisearch.com>	2022-09-21 09:14:46 +00:00
mlemesle	d406fe901b	Pass config.toml keys to snake_case	2022-09-21 10:55:16 +02:00
Loïc Lecrenier	513a38f07b	Remove LTO in release profile Since we can't enable it in Meilisearch, there is no point in having it enabled in milli	2022-09-21 10:44:33 +02:00
bors[bot]	e1e025c319	Merge #641 641: Remove `helpers` crate r=Kerollmops a=loiclec # Pull Request ## What does this PR do? Remove the `helpers` crates, because (I think) we don't use it. This should have been part of https://github.com/meilisearch/milli/pull/636 , but I forgot about it then :) Co-authored-by: Loïc Lecrenier <loic@meilisearch.com>	2022-09-21 08:36:05 +00:00
Loïc Lecrenier	b6fe6838d3	Remove `helpers` crate	2022-09-21 10:25:36 +02:00
mlemesle	4dfae44478	Apply PR review comments	2022-09-19 18:16:28 +02:00
Jakub Jirutka	935f18efcf	Allow building without specialized tokenizations (Some of) these specialized tokenizations include huge dictionaries that currently account for 90% (!) of the meilisearch binary size. This commit adds chinese, hebrew, japanese, and thai feature flags that are propagated via milli down to the charabia crate. To keep it backward compatible, they are enabled by default. Related to meilisearch/milli#632	2022-09-14 21:16:34 +02:00
Jakub Jirutka	5b57114771	Bump milli from 0.33.0 to 0.33.4	2022-09-14 20:52:11 +02:00
Clémentine Urquizar - curqui	c2ab7a7939	Update config.yml	2022-09-14 14:40:36 +02:00
bors[bot]	d94339a858	Merge #636 636: Remove unused `infos`, `http-ui`, and `milli/fuzz`, crates r=ManyTheFish a=loiclec We haven't used the `infos/`, `http-ui/` and `milli/fuzz/` crates in a long time. They are not properly maintained and probably do not work correctly anymore. This PR removes these crates entirely from the workspace to reduce the amount of code we need to maintain. Co-authored-by: Loïc Lecrenier <loic@meilisearch.com>	2022-09-14 12:39:57 +00:00
bors[bot]	15d478cf4d	Merge #635 635: Use an unstable algorithm for `grenad::Sorter` when possible r=Kerollmops a=loiclec # Pull Request ## What does this PR do? Use an unstable algorithm to sort the internal vector used by `grenad::Sorter` whenever possible to speed up indexing. In practice, every time the merge function creates a `RoaringBitmap`, we use an unstable sort. For every other merge function, such as `keep_first`, `keep_last`, etc., a stable sort is used. Co-authored-by: Loïc Lecrenier <loic@meilisearch.com>	2022-09-14 12:00:52 +00:00
Loïc Lecrenier	add96f921b	Remove unused infos/ http-ui/ and fuzz/ crates	2022-09-14 06:55:01 +02:00
bors[bot]	fa315352da	Merge #2770 2770: Update milli 0.33.4 r=Kerollmops a=curquiza Fixes https://github.com/meilisearch/meilisearch/issues/2764 Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-09-13 16:07:06 +00:00
Clémentine Urquizar	268d59ccb1	Update milli version to v0.33.4	2022-09-13 18:01:09 +02:00
bors[bot]	4fc6331cb6	Merge #638 638: Update version for the next release (v0.33.4) in Cargo.toml files r=curquiza a=meili-bot ⚠️ This PR is automatically generated. Check the new version is the expected one before merging. Co-authored-by: curquiza <curquiza@users.noreply.github.com>	2022-09-13 13:56:53 +00:00
curquiza	753e76d451	Update version for the next release (v0.33.4) in Cargo.toml files	2022-09-13 13:55:50 +00:00
Loïc Lecrenier	3794962330	Use an unstable algorithm for grenad::Sorter when possible	2022-09-13 14:49:53 +02:00
bors[bot]	2865b063ad	Merge #637 637: We avoid skipping errors in the indexing pipeline r=ManyTheFish a=Kerollmops This PR is related to https://github.com/meilisearch/meilisearch/issues/2764 and should fix it when merged into Meilisearch. Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-09-13 12:12:05 +00:00
Kerollmops	d4d7c9d577	We avoid skipping errors in the indexing pipeline	2022-09-13 14:03:00 +02:00
bors[bot]	5901d4e407	Merge #2768 2768: Update patch versions to remove CVE r=Kerollmops a=curquiza Trying to fix CVE we have with [synchronoise](https://github.com/QuietMisdreavus/synchronoise) crate Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-09-12 12:47:59 +00:00
Clémentine Urquizar	aabd67a9fa	Update patch version to remove CVE	2022-09-12 14:36:45 +02:00
mlemesle	a690ace36e	Add example config.toml with default values	2022-09-09 09:37:23 +02:00
bors[bot]	f8697075ea	Merge #632 632: Make charabia default feature optional r=ManyTheFish a=vincent-herlemont # Pull Request ## What does this PR do? Fixes [#627](https://github.com/meilisearch/milli/issues/627#issuecomment-1239769122) Thank you so much for contributing to Meilisearch! Co-authored-by: Vincent Herlemont <vincent@herlemont.fr>	2022-09-08 14:33:26 +00:00
bors[bot]	7cd0aea1d3	Merge #633 633: Upgrade ubuntu-18.04 to 20.04 r=Kerollmops a=curquiza Ubuntu-18.04 is going to be deprecated by GitHub https://github.com/actions/runner-images/issues/6002 Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-09-08 14:08:28 +00:00
Clémentine Urquizar	69b2d31b71	Upgrade ubuntu-18.04 to 20.04	2022-09-08 14:58:06 +02:00
Vincent Herlemont	8cd5200f48	Make charabia languages configurable	2022-09-08 12:21:43 +02:00
mlemesle	579fa3f1ad	Remove unnecessary println	2022-09-08 11:05:52 +02:00
bors[bot]	99b45a7820	Merge #631 631: Revert "Remove Bors required test for Windows" r=Kerollmops a=curquiza Reverts meilisearch/milli#612 Because the issue does not seem to be there! Closes https://github.com/meilisearch/milli/issues/614 Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com>	2022-09-07 21:07:44 +00:00
bors[bot]	3fd6af25f9	Merge #2759 2759: Bump milli to 0.33.3 r=Kerollmops a=Kerollmops This PR fixes #2743. Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-09-07 20:55:08 +00:00
Vincent Herlemont	5e07ea79c2	Make charabia default feature optional	2022-09-07 20:54:31 +02:00
mlemesle	7f267ec4be	Fix clippy	2022-09-07 20:22:49 +02:00
Clémentine Urquizar - curqui	3af3d3f7d9	Revert "Remove Bors required test for Windows"	2022-09-07 18:36:10 +02:00
Kerollmops	441492f1c8	Bump milli to v0.33.3	2022-09-07 18:23:49 +02:00
bors[bot]	528a4721c1	Merge #2727 2727: Don't panic when the error length is slightly over 100 r=Kerollmops a=onyxcherry # Pull Request ## What does this PR do? Fixes PR #2207 as [the last commit](`7ece7a9d9e`) has changed number of the characters at the end to leave in place from `50` to `85` but the lower limit of a string length wasn't changed. Therefore, any data (e.g. example string from issue #2680) was causing `meilisearch` to panic. So I simply raised the minimum value from `100` to `135` (`50 + 85`) to ensure that `replace_range()` won't panic due to an inverted range. At the same time I am in favor of the `85` value which was changed in the `@CNLHC's` last commit. ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing ~issue~ pull request? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Tomasz Wiśniewski <tomasz@wisniewski.app>	2022-09-07 16:16:43 +00:00
mlemesle	5a4f1508d0	Add documentation	2022-09-07 18:16:33 +02:00
Tomasz Wiśniewski	d1a30df23d	Remove unneeded prints, format	2022-09-07 18:05:55 +02:00
bors[bot]	549fa12d5a	Merge #629 629: Update version for the next release (v0.33.3) in Cargo.toml files r=curquiza a=meili-bot ⚠️ This PR is automatically generated. Check the new version is the expected one before merging. Co-authored-by: curquiza <curquiza@users.noreply.github.com>	2022-09-07 15:55:04 +00:00
bors[bot]	92b0c51bfe	Merge #2755 2755: Update mini-dashboard to v0.2.2 r=Kerollmops a=mdubus # Pull Request ## What does this PR do? Fixes #2716 ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Morgane Dubus <30866152+mdubus@users.noreply.github.com>	2022-09-07 15:53:04 +00:00
curquiza	077dcd2002	Update version for the next release (v0.33.3) in Cargo.toml files	2022-09-07 15:48:53 +00:00
mlemesle	135499f398	Extract new env vars to const	2022-09-07 17:47:15 +02:00
bors[bot]	b3ffcb2d97	Merge #2758 2758: Update ubuntu-18.04 to 20.04 r=Kerollmops a=curquiza Trying to avoid CI failure by updating ubuntu machines Commit already available on main, so for v0.30.0 https://github.com/meilisearch/meilisearch/pull/2719 Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-09-07 15:27:19 +00:00
Clémentine Urquizar	5cbd047989	Update ubuntu-18.04 to 20.04	2022-09-07 17:24:35 +02:00
mlemesle	ef3fa92536	Refactor default values for clap and serde	2022-09-07 16:58:03 +02:00
mlemesle	6520d3c474	Refactor build method and flag	2022-09-07 16:09:00 +02:00
mlemesle	403226a029	Add support for config file	2022-09-07 16:09:00 +02:00
bors[bot]	2907928d93	Merge #628 628: Make sure that long words are ignored r=ManyTheFish a=Kerollmops This PR is related to https://github.com/meilisearch/meilisearch/issues/2743 and is fixing it. Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-09-07 13:04:59 +00:00
Kerollmops	fe3973a51c	Make sure that long words are correctly skipped	2022-09-07 15:03:32 +02:00
Kerollmops	c83c3cd796	Add a test to make sure that long words are correctly skipped	2022-09-07 14:12:36 +02:00
Morgane Dubus	07f45251e9	Update mini-dashboard to v0.2.2	2022-09-07 11:09:12 +02:00
bors[bot]	b9539c59f3	Merge #625 #626 625: Bump actions/checkout from 2 to 3 r=curquiza a=dependabot[bot] Bumps [actions/checkout](https://github.com/actions/checkout) from 2 to 3. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/actions/checkout/releases">actions/checkout's releases</a>.</em></p> <blockquote> <h2>v3.0.0</h2> <ul> <li>Updated to the node16 runtime by default <ul> <li>This requires a minimum <a href="https://github.com/actions/runner/releases/tag/v2.285.0">Actions Runner</a> version of v2.285.0 to run, which is by default available in GHES 3.4 or later.</li> </ul> </li> </ul> <h2>v2.4.2</h2> <h2>What's Changed</h2> <ul> <li>Add set-safe-directory input to allow customers to take control. (<a href="https://github-redirect.dependabot.com/actions/checkout/issues/770">#770</a>) by <a href="https://github.com/TingluoHuang"><code>`@TingluoHuang</code></a>` in <a href="https://github-redirect.dependabot.com/actions/checkout/pull/776">actions/checkout#776</a></li> <li>Prepare changelog for v2.4.2. by <a href="https://github.com/TingluoHuang"><code>`@TingluoHuang</code></a>` in <a href="https://github-redirect.dependabot.com/actions/checkout/pull/778">actions/checkout#778</a></li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/actions/checkout/compare/v2...v2.4.2">https://github.com/actions/checkout/compare/v2...v2.4.2</a></p> <h2>v2.4.1</h2> <ul> <li>Fixed an issue where checkout failed to run in container jobs due to the new git setting <code>safe.directory</code></li> </ul> <h2>v2.4.0</h2> <ul> <li>Convert SSH URLs like <code>org-<ORG_ID>`@github.com:</code>` to <code>https://github.com/</code> - <a href="https://github-redirect.dependabot.com/actions/checkout/pull/621">pr</a></li> </ul> <h2>v2.3.5</h2> <p>Update dependencies</p> <h2>v2.3.4</h2> <ul> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/379">Add missing <code>await</code>s</a></li> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/360">Swap to Environment Files</a></li> </ul> <h2>v2.3.3</h2> <ul> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/345">Remove Unneeded commit information from build logs</a></li> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/326">Add Licensed to verify third party dependencies</a></li> </ul> <h2>v2.3.2</h2> <p><a href="https://github-redirect.dependabot.com/actions/checkout/pull/320">Add Third Party License Information to Dist Files</a></p> <h2>v2.3.1</h2> <p><a href="https://github-redirect.dependabot.com/actions/checkout/pull/284">Fix default branch resolution for .wiki and when using SSH</a></p> <h2>v2.3.0</h2> <p><a href="https://github-redirect.dependabot.com/actions/checkout/pull/278">Fallback to the default branch</a></p> <h2>v2.2.0</h2> <p><a href="https://github-redirect.dependabot.com/actions/checkout/pull/258">Fetch all history for all tags and branches when fetch-depth=0</a></p> <h2>v2.1.1</h2> <p>Changes to support GHES (<a href="https://github-redirect.dependabot.com/actions/checkout/pull/236">here</a> and <a href="https://github-redirect.dependabot.com/actions/checkout/pull/248">here</a>)</p> <h2>v2.1.0</h2> <ul> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/191">Group output</a></li> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/199">Changes to support GHES alpha release</a></li> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/184">Persist core.sshCommand for submodules</a></li> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/163">Add support ssh</a></li> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/179">Convert submodule SSH URL to HTTPS, when not using SSH</a></li> </ul> <!-- raw HTML omitted --> </blockquote> <p>... (truncated)</p> </details> <details> <summary>Changelog</summary> <p><em>Sourced from <a href="https://github.com/actions/checkout/blob/main/CHANGELOG.md">actions/checkout's changelog</a>.</em></p> <blockquote> <h1>Changelog</h1> <h2>v3.0.2</h2> <ul> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/770">Add input <code>set-safe-directory</code></a></li> </ul> <h2>v3.0.1</h2> <ul> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/762">Fixed an issue where checkout failed to run in container jobs due to the new git setting <code>safe.directory</code></a></li> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/744">Bumped various npm package versions</a></li> </ul> <h2>v3.0.0</h2> <ul> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/689">Update to node 16</a></li> </ul> <h2>v2.3.1</h2> <ul> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/284">Fix default branch resolution for .wiki and when using SSH</a></li> </ul> <h2>v2.3.0</h2> <ul> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/278">Fallback to the default branch</a></li> </ul> <h2>v2.2.0</h2> <ul> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/258">Fetch all history for all tags and branches when fetch-depth=0</a></li> </ul> <h2>v2.1.1</h2> <ul> <li>Changes to support GHES (<a href="https://github-redirect.dependabot.com/actions/checkout/pull/236">here</a> and <a href="https://github-redirect.dependabot.com/actions/checkout/pull/248">here</a>)</li> </ul> <h2>v2.1.0</h2> <ul> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/191">Group output</a></li> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/199">Changes to support GHES alpha release</a></li> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/184">Persist core.sshCommand for submodules</a></li> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/163">Add support ssh</a></li> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/179">Convert submodule SSH URL to HTTPS, when not using SSH</a></li> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/157">Add submodule support</a></li> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/144">Follow proxy settings</a></li> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/141">Fix ref for pr closed event when a pr is merged</a></li> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/128">Fix issue checking detached when git less than 2.22</a></li> </ul> <h2>v2.0.0</h2> <ul> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/108">Do not pass cred on command line</a></li> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/107">Add input persist-credentials</a></li> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/104">Fallback to REST API to download repo</a></li> </ul> <h2>v2 (beta)</h2> <ul> <li>Improved fetch performance</li> </ul> <!-- raw HTML omitted --> </blockquote> <p>... (truncated)</p> </details> <details> <summary>Commits</summary> <ul> <li><a href="`2541b1294d`"><code>2541b12</code></a> Prepare changelog for v3.0.2. (<a href="https://github-redirect.dependabot.com/actions/checkout/issues/777">#777</a>)</li> <li><a href="`0ffe6f9c55`"><code>0ffe6f9</code></a> Add set-safe-directory input to allow customers to take control. (<a href="https://github-redirect.dependabot.com/actions/checkout/issues/770">#770</a>)</li> <li><a href="`dcd71f6466`"><code>dcd71f6</code></a> Enforce safe directory (<a href="https://github-redirect.dependabot.com/actions/checkout/issues/762">#762</a>)</li> <li><a href="`add3486cc3`"><code>add3486</code></a> Patch to fix the dependbot alert. (<a href="https://github-redirect.dependabot.com/actions/checkout/issues/744">#744</a>)</li> <li><a href="`5126516654`"><code>5126516</code></a> Bump minimist from 1.2.5 to 1.2.6 (<a href="https://github-redirect.dependabot.com/actions/checkout/issues/741">#741</a>)</li> <li><a href="`d50f8ea767`"><code>d50f8ea</code></a> Add v3.0 release information to changelog (<a href="https://github-redirect.dependabot.com/actions/checkout/issues/740">#740</a>)</li> <li><a href="`2d1c1198e7`"><code>2d1c119</code></a> update test workflows to checkout v3 (<a href="https://github-redirect.dependabot.com/actions/checkout/issues/709">#709</a>)</li> <li><a href="`a12a3943b4`"><code>a12a394</code></a> update readme for v3 (<a href="https://github-redirect.dependabot.com/actions/checkout/issues/708">#708</a>)</li> <li><a href="`8f9e05e482`"><code>8f9e05e</code></a> Update to node 16 (<a href="https://github-redirect.dependabot.com/actions/checkout/issues/689">#689</a>)</li> <li>See full diff in <a href="https://github.com/actions/checkout/compare/v2...v3">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=actions/checkout&package-manager=github_actions&previous-version=2&new-version=3)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) You can trigger a rebase of this PR by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) </details> 626: Bump yogevbd/enforce-label-action from 2.1.0 to 2.2.2 r=curquiza a=dependabot[bot] Bumps [yogevbd/enforce-label-action](https://github.com/yogevbd/enforce-label-action) from 2.1.0 to 2.2.2. <details> <summary>Commits</summary> <ul> <li><a href="`a3c219da6b`"><code>a3c219d</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/yogevbd/enforce-label-action/issues/26">#26</a> from yogevbd/test1</li> <li><a href="`8279da6fd9`"><code>8279da6</code></a> Update enforce-labels.yml</li> <li><a href="`0c6f806593`"><code>0c6f806</code></a> Update package.json</li> <li><a href="`2e6b1550e4`"><code>2e6b155</code></a> lock <code>`@actions/http-client</code></li>` <li><a href="`732db2ff3a`"><code>732db2f</code></a> test</li> <li><a href="`e662799851`"><code>e662799</code></a> Update package.json</li> <li><a href="`f467829919`"><code>f467829</code></a> Update enforce-labels.yml</li> <li><a href="`00ff95bb80`"><code>00ff95b</code></a> Update package.json</li> <li><a href="`de4244ae68`"><code>de4244a</code></a> Update action.yml</li> <li><a href="`9f40e51d60`"><code>9f40e51</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/yogevbd/enforce-label-action/issues/25">#25</a> from dominikmeyersap/patch-1</li> <li>Additional commits viewable in <a href="https://github.com/yogevbd/enforce-label-action/compare/2.1.0...2.2.2">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=yogevbd/enforce-label-action&package-manager=github_actions&previous-version=2.1.0&new-version=2.2.2)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) You can trigger a rebase of this PR by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) </details> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2022-09-06 16:49:30 +00:00
bors[bot]	f2b140d3d7	Merge #624 624: Bump Swatinem/rust-cache from 1.3.0 to 2.0.0 r=curquiza a=dependabot[bot] Bumps [Swatinem/rust-cache](https://github.com/Swatinem/rust-cache) from 1.3.0 to 2.0.0. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/Swatinem/rust-cache/releases">Swatinem/rust-cache's releases</a>.</em></p> <blockquote> <h2>v2.0.0</h2> <ul> <li>The action code was refactored to allow for caching multiple workspaces and different <code>target</code> directory layouts.</li> <li>The <code>working-directory</code> and <code>target-dir</code> input options were replaced by a single <code>workspaces</code> option that has the form of <code>$workspace -> $target</code>.</li> <li>Support for considering <code>env-vars</code> as part of the cache key.</li> <li>The <code>sharedKey</code> input option was renamed to <code>shared-key</code> for consistency.</li> </ul> <h2>v1.4.0</h2> <ul> <li>Clean both debug and release target directories.</li> </ul> </blockquote> </details> <details> <summary>Changelog</summary> <p><em>Sourced from <a href="https://github.com/Swatinem/rust-cache/blob/master/CHANGELOG.md">Swatinem/rust-cache's changelog</a>.</em></p> <blockquote> <h2>2.0.0</h2> <ul> <li>The action code was refactored to allow for caching multiple workspaces and different <code>target</code> directory layouts.</li> <li>The <code>working-directory</code> and <code>target-dir</code> input options were replaced by a single <code>workspaces</code> option that has the form of <code>$workspace -> $target</code>.</li> <li>Support for considering <code>env-vars</code> as part of the cache key.</li> <li>The <code>sharedKey</code> input option was renamed to <code>shared-key</code> for consistency.</li> </ul> <h2>1.4.0</h2> <ul> <li>Clean both <code>debug</code> and <code>release</code> target directories.</li> </ul> </blockquote> </details> <details> <summary>Commits</summary> <ul> <li><a href="`6720f05bc4`"><code>6720f05</code></a> 2.0.0</li> <li><a href="`5733786579`"><code>5733786</code></a> rebuild</li> <li><a href="`622616010e`"><code>6226160</code></a> prepare v2</li> <li><a href="`0497f9301f`"><code>0497f93</code></a> improve registry cleanpu</li> <li><a href="`7b8626742a`"><code>7b86267</code></a> update registry cleaning</li> <li><a href="`911d8e9e55`"><code>911d8e9</code></a> test sparse registry</li> <li><a href="`875be5ce2d`"><code>875be5c</code></a> bump cache</li> <li><a href="`07a2ee71bc`"><code>07a2ee7</code></a> lol, dependency check was reversed</li> <li><a href="`7c190ef171`"><code>7c190ef</code></a> fix actual test code ;-)</li> <li><a href="`fffd6895b2`"><code>fffd689</code></a> add some more tests</li> <li>Additional commits viewable in <a href="https://github.com/Swatinem/rust-cache/compare/v1.3.0...v2.0.0">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=Swatinem/rust-cache&package-manager=github_actions&previous-version=1.3.0&new-version=2.0.0)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) You can trigger a rebase of this PR by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) </details> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2022-09-06 16:12:55 +00:00
dependabot[bot]	e3400a05d3	Bump yogevbd/enforce-label-action from 2.1.0 to 2.2.2 Bumps [yogevbd/enforce-label-action](https://github.com/yogevbd/enforce-label-action) from 2.1.0 to 2.2.2. - [Release notes](https://github.com/yogevbd/enforce-label-action/releases) - [Commits](https://github.com/yogevbd/enforce-label-action/compare/2.1.0...2.2.2) --- updated-dependencies: - dependency-name: yogevbd/enforce-label-action dependency-type: direct:production update-type: version-update:semver-minor ... Signed-off-by: dependabot[bot] <support@github.com>	2022-09-06 16:08:54 +00:00
dependabot[bot]	b308463022	Bump actions/checkout from 2 to 3 Bumps [actions/checkout](https://github.com/actions/checkout) from 2 to 3. - [Release notes](https://github.com/actions/checkout/releases) - [Changelog](https://github.com/actions/checkout/blob/main/CHANGELOG.md) - [Commits](https://github.com/actions/checkout/compare/v2...v3) --- updated-dependencies: - dependency-name: actions/checkout dependency-type: direct:production update-type: version-update:semver-major ... Signed-off-by: dependabot[bot] <support@github.com>	2022-09-06 16:08:51 +00:00
dependabot[bot]	5e85059a71	Bump Swatinem/rust-cache from 1.3.0 to 2.0.0 Bumps [Swatinem/rust-cache](https://github.com/Swatinem/rust-cache) from 1.3.0 to 2.0.0. - [Release notes](https://github.com/Swatinem/rust-cache/releases) - [Changelog](https://github.com/Swatinem/rust-cache/blob/master/CHANGELOG.md) - [Commits](https://github.com/Swatinem/rust-cache/compare/v1.3.0...v2.0.0) --- updated-dependencies: - dependency-name: Swatinem/rust-cache dependency-type: direct:production update-type: version-update:semver-major ... Signed-off-by: dependabot[bot] <support@github.com>	2022-09-06 16:08:48 +00:00
bors[bot]	9e661f2cb9	Merge #623 623: Add dependabot for GHA r=Kerollmops a=curquiza Same as we added in Meilisearch. Only runs once a month. https://github.com/meilisearch/meilisearch/blob/main/.github/dependabot.yml Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-09-06 15:56:28 +00:00
Clémentine Urquizar	44192d754f	Add dependabot for GHA	2022-09-06 17:54:05 +02:00
bors[bot]	1fa851a8d0	Merge #622 622: Minor fixes in the just added update-version CI r=ManyTheFish a=curquiza These fixes are minor, and do not prevent us to use the current CI Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-09-06 13:14:23 +00:00
bors[bot]	3a62e97c9f	Merge #2744 2744: Minor fixes in the just added update-version CI r=ManyTheFish a=curquiza These fixes does not prevent us to use the current CI Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-09-06 13:14:06 +00:00
Many the fish	37dc6537c3	Fix api keys bugs (#2734 ) * Add some tests * Disallow index creation when API key doesn't havec explicitelly the right on the creating index * Fix lazy index creation with `indexes.*` action	2022-09-06 15:13:09 +02:00
Clémentine Urquizar	55aa83d75a	Minor fixes in the just added update-version CI	2022-09-05 16:41:01 +02:00
Clémentine Urquizar	61abc61a69	Minor fixes in the just added update-version CI	2022-09-05 16:01:32 +02:00
bors[bot]	0c962f6a76	Merge #2740 2740: Update checkout v2 to v3 in CI manifests and use a unique GitHub PAT r=Kerollmops a=curquiza Upgrade the missing checkout v2 into v3 Probably a bad merge conflicts that make them removed when merging `stable` into `main` after v0.28.0 release. Also, use `MEILI_BOT_GH_PAT` instead of `PUBLISH_TOKEN` and the default github token, which allow us to remove useless GitHub secrets (once this PR is merged and v0.29.0 is release because `PUBLISH_TOKEN` is still used on `release-v0.29.0`) Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-09-05 12:14:40 +00:00
bors[bot]	2e0221cfc6	Merge #2739 2739: Add CI manifest to automate some steps when closing/creating a Milestone r=Kerollmops a=curquiza For each Milestone created (not opened!), and if the release is NOT a patch release (only the patch changed) - the roadmap issue is created, see the [template](https://github.com/meilisearch/core-team/blob/main/issue-templates/roadmap-issue.md) - the changelog issue is created, see the [template](https://github.com/meilisearch/core-team/blob/main/issue-templates/changelog-issue.md) For each Milestone closed - the `vX.Y.Z` label is created - this label is applied to all issues/PRs in the Milestone Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-09-05 11:42:10 +00:00
bors[bot]	2d34b239ab	Merge #2738 2738: Add missing env vars for dumps and snapshots features r=irevoire a=gmourier # Pull Request ## What does this PR do? Fixes #2721 ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Guillaume Mourier <guillaume@meilisearch.com>	2022-09-05 11:12:09 +00:00
bors[bot]	55853815a5	Merge #2741 2741: Add CI to update the Meilisearch version in Cargo.toml files r=ManyTheFish a=curquiza Add a CI we can trigger manually to create a PR updating the Meilisearch version The next step is to create a Slack bot that will trigger this CI In the meantime, we can trigger this CI manually in the [Actions tab](https://github.com/meilisearch/milli/actions) The `MEILI_BOT_GH_PAT` secrets has been added to the organization level, and is accessible for the following repositories (so far): Meilisearch, Milli and Charabia Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-09-05 08:32:43 +00:00
bors[bot]	efee0e3f43	Merge #621 621: Add CI to update the Milli version r=ManyTheFish a=curquiza Add a CI we can trigger manually to create a PR updating the Milli version The next step is to create a Slack bot that will trigger this CI In the meantime, we can trigger this CI manually in the [Actions tab](https://github.com/meilisearch/milli/actions) The `MEILI_BOT_GH_PAT` secrets has been added to the organization level, and is accessible for the following repositories (so far): Meilisearch, Milli and Charabia Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-09-05 08:31:48 +00:00
Clémentine Urquizar	b897ce8dfa	Add CI to update th Meilisearch version in Cargo.toml files	2022-09-04 11:57:23 +02:00
Clémentine Urquizar	0639b14906	Add CI to update the Milli version	2022-09-04 11:49:50 +02:00
Clémentine Urquizar	5f5b787483	Use meili-bot PAT everywhere	2022-09-04 11:32:22 +02:00
Clémentine Urquizar	d0aa8042e2	Fix job names	2022-09-03 17:53:37 +02:00
Clémentine Urquizar	9cb1e4af5c	Rename workflow file	2022-09-03 17:50:24 +02:00
Clémentine Urquizar	cc09aa8868	Refacto env var	2022-09-02 17:07:23 +02:00
Clémentine Urquizar	2eca723a91	Update checkout v2 to v3 in CI manifests	2022-09-02 16:45:32 +02:00
Clémentine Urquizar	ae14567f97	Add CI manifest to automate some step of the release management when creating/closing a Milestone	2022-09-02 16:28:58 +02:00
Guillaume Mourier	d0f1054f5c	chore: cargo fmt	2022-09-01 22:37:07 +02:00
Guillaume Mourier	3878c289df	feat: add missing env var for dumps and snapshots feature	2022-09-01 22:34:20 +02:00
Tomasz Wiśniewski	d1b3642923	Extract input to trim lengths to variables	2022-09-01 20:50:11 +02:00
bors[bot]	f45ddff312	Merge #2733 2733: Move if conditions to the steps and not to the whole job r=Kerollmops a=curquiza Follows https://github.com/meilisearch/meilisearch/pull/2726 Fixes the CI. Moves the `if:` condition to the steps and not the whole `check-version` jobs, otherwise, the following jobs checking the compilation cannot run. See: https://github.com/meilisearch/meilisearch/actions/runs/2968543407 Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-09-01 11:49:40 +00:00
Clémentine Urquizar	8f3b590436	Move if conditions to the steps and not to the whole job	2022-09-01 13:34:34 +02:00
bors[bot]	4e37427de8	Merge #2732 2732: Update milli v0.33.2 r=Kerollmops a=ManyTheFish closes #2722 ⚠️ : merging into [release-v0.29.0](https://github.com/meilisearch/meilisearch/tree/release-v0.29.0) Co-authored-by: ManyTheFish <many@meilisearch.com>	2022-09-01 11:18:11 +00:00
ManyTheFish	50434d35d0	Update milli v0.33.2	2022-09-01 13:15:05 +02:00
bors[bot]	f7c352a32d	Merge #620 620: Fix word criterion r=Kerollmops a=ManyTheFish related to https://github.com/meilisearch/meilisearch/issues/2722 - fix the word strategy bug - update milli version to v0.33.2 Co-authored-by: ManyTheFish <many@meilisearch.com>	2022-09-01 10:14:35 +00:00
ManyTheFish	bf750e45a1	Fix word removal issue	2022-09-01 12:10:47 +02:00
ManyTheFish	a38608fe59	Add test mixing phrased and no-phrased words	2022-09-01 12:02:10 +02:00
ManyTheFish	97a04887a3	Update version for next release (v0.33.2) in Cargo.toml	2022-09-01 11:47:23 +02:00
Tomasz Wiśniewski	60792eebcf	Fix #2207 - do not panic when the error message length is between 100 and 135	2022-08-31 19:29:53 +02:00
bors[bot]	4fbc8c705f	Merge #2726 2726: Add dry run for publishing binaries: check the compilation works r=Kerollmops a=curquiza To avoid realizing the compilation of one type of binary does not work during the release, I create a dry run for binary compilation every day at 2am 😇 See the problem we had recently because missing this CI: https://github.com/meilisearch/meilisearch/issues/2718 Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-08-31 15:35:05 +00:00
Clémentine Urquizar	80b1f3e830	Add dry run for publishing binaries: check the compilation works	2022-08-31 17:28:42 +02:00
bors[bot]	e315547ffc	Merge #2724 2724: Make the document addition done log to appear once indexing is over r=curquiza a=evpeople # Pull Request ## What does this PR do? Fixes #2703 <!-- Please link the issue you're trying to fix with this PR, if none then please create an issue first. --> ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: evpeople <hangcaihui@gmail.com>	2022-08-31 13:13:36 +00:00
bors[bot]	17d020e996	Merge #618 618: Update version for next release (v0.33.1) in Cargo.toml r=Kerollmops a=curquiza No breaking for this release Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-08-31 10:43:45 +00:00
Clémentine Urquizar	c3363706c5	Update version for next release (v0.33.1) in Cargo.toml	2022-08-31 11:37:27 +02:00
bors[bot]	2c2f3d38cc	Merge #617 617: Accept integers as document ids again r=irevoire a=Kerollmops This PR is related to https://github.com/meilisearch/meilisearch/issues/2723 and will fix when this PR will be merged, a new release deployed and used in Meilisearch itself. This PR makes the indexer to try to parse the values of the fields identified as numbers i.e. `id:number` as integer first then as float if it fails. Co-authored-by: Clément Renault <clement@meilisearch.com>	2022-08-31 09:25:17 +00:00
Clément Renault	7f92116b51	Accept again integers as document ids	2022-08-31 10:56:39 +02:00
evpeople	833ade80a6	cargo update	2022-08-31 13:58:53 +08:00
evpeople	f117c90c46	remove the intermediate addition variable	2022-08-30 21:49:34 +08:00
bors[bot]	3efd06168f	Merge #2719 2719: Update ubuntu-18.04 to 20.04 r=Kerollmops a=curquiza ubuntu-18.04 is deprecated <img width="1566" alt="Capture d’écran 2022-08-30 à 12 04 25" src="https://user-images.githubusercontent.com/20380692/187409661-5ac9c545-a737-47f5-ad2f-0369114be11a.png"> https://github.com/meilisearch/meilisearch/actions/runs/2949985452 Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-08-30 13:42:09 +00:00
Clémentine Urquizar	07bb32b622	Update ubuntu-18.04 to 20.04	2022-08-29 18:26:01 +02:00
bors[bot]	1131400694	Merge #2717 2717: Disable LTO due to compilation failures on some platforms r=curquiza a=loiclec Meilisearch fails to compile on aarch64 Linux due to a linker error ( https://github.com/meilisearch/meilisearch/runs/8072616457?check_suite_focus=true ). This is probably caused by link-time-optimisation (LTO). Since it is not possible to modify a profile based on the target triple, this PR deactivates LTO completely for all platforms. In the future, we might want to create different custom profiles, such as: ```toml [profile.release-lto] inherits = "release" lto = "thin" ``` and compile Meilisearch using `cargo build --profile release-lto` on the platforms that can support it. Co-authored-by: Loïc Lecrenier <loic@meilisearch.com>	2022-08-29 16:04:34 +00:00
Loïc Lecrenier	9a789daa58	Disable LTO due to compilation failures on some platforms (aarch64 linux)	2022-08-29 17:21:08 +02:00
bors[bot]	0826aa35e2	Merge #2713 2713: Move prometheus behind a feature flag r=Kerollmops a=irevoire We decided we wanted to continue working on this feature before making it public. Co-authored-by: Tamo <tamo@meilisearch.com>	2022-08-29 12:38:42 +00:00
Tamo	6aa3ad6b6c	move prometheus behind a feature flag	2022-08-29 14:36:59 +02:00
evpeople	b774adfbf7	The "document addition done" appear once the indexation is over now.	2022-08-27 21:50:10 +08:00
bors[bot]	47a1aa69f0	Merge #2702 2702: Add link to the main image r=curquiza a=brunoocasali I have wrapped the image with a `<a>` link, and it seems to be working fine, WDYT? Co-authored-by: Bruno Casali <brunoocasali@gmail.com> Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com>	2022-08-24 16:54:08 +00:00
Clémentine Urquizar - curqui	b3d89da74d	Update README.md	2022-08-24 18:52:34 +02:00
Bruno Casali	0798170a64	Add link to the main image	2022-08-24 13:47:09 -03:00
bors[bot]	446dfccc8c	Merge #2504 #2697 2504: New README 🌟 r=curquiza a=curquiza ⚠️ Please do not only look at the Markdown but also how the GitHub renders the README 😇 👉 👉 [Rendered](https://github.com/meilisearch/meilisearch/blob/new-readme/README.md) 👈 👈 2697: Accept an environment variable to enable the metrics route r=ManyTheFish a=Kerollmops With the PR Meilisearch is able to accept the `MEILI_ENABLE_METRICS_ROUTE` environment variable to enable the newly introduces metrics route. Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com> Co-authored-by: Clément Renault <clement@meilisearch.com>	2022-08-24 15:41:44 +00:00
Clémentine Urquizar	43175cfb78	New README version Update README.md Co-authored-by: gui machiavelli <hey@guimachiavelli.com> Update according to guigui reuqest Add demo link Update README.md Co-authored-by: gui machiavelli <hey@guimachiavelli.com> Update README.md Co-authored-by: gui machiavelli <hey@guimachiavelli.com> Update README.md Co-authored-by: gui machiavelli <hey@guimachiavelli.com> Update README.md Co-authored-by: gui machiavelli <hey@guimachiavelli.com> Update README.md Co-authored-by: gui machiavelli <hey@guimachiavelli.com> Update README.md Co-authored-by: CaroFG <48251481+CaroFG@users.noreply.github.com> Update README.md Co-authored-by: gui machiavelli <hey@guimachiavelli.com> Update README.md Co-authored-by: gui machiavelli <hey@guimachiavelli.com> Update README.md Co-authored-by: gui machiavelli <hey@guimachiavelli.com> Update README.md Co-authored-by: gui machiavelli <hey@guimachiavelli.com> Put sentence in bold	2022-08-24 17:16:29 +02:00
bors[bot]	a1bb49c351	Merge #2696 2696: Add the new `metrics.get` and `metrics.all` actions rights r=Kerollmops a=Kerollmops Follow the specification and add the new `metrics.get` and `metrics.all` actions, making the `/metrics` route only accessible with those rights. Co-authored-by: Clément Renault <clement@meilisearch.com>	2022-08-24 15:08:53 +00:00
Clément Renault	bebd76064a	Add test for the rights of /metrics route	2022-08-24 17:03:43 +02:00
Clément Renault	f0b2ac6efb	metrics.all must define metrics.get	2022-08-24 17:03:30 +02:00
Clément Renault	08d86e33ca	Accept an env variable to enable the metrics route	2022-08-24 16:39:56 +02:00
Clément Renault	2c2efc7ab6	Remove the hand written numbers of the actions rights	2022-08-24 16:33:12 +02:00
Clément Renault	381df43be4	Change the metrics route API access rights	2022-08-24 16:28:33 +02:00
bors[bot]	f87ebfe477	Merge #2692 2692: Slight changes for prometheus metrics r=Kerollmops a=gmourier # Pull Request ## What does this PR do? - Replace "MeiliSearch" with "Meilisearch" - Brings some consistency between rust identifier and exposed metrics names - Add suffix describing unit, in plural form. e.g `MEILISEARCH_DB_SIZE_BYTES` (https://prometheus.io/docs/practices/naming/#metric-names) - Update dashboard.json Co-authored-by: Guillaume Mourier <guillaume@meilisearch.com>	2022-08-24 10:12:24 +00:00
bors[bot]	c445334070	Merge #2636 2636: Upgrade milli to v0.33.0 r=Kerollmops a=ManyTheFish # Summary - Update milli to v0.33.0 - Classify the new InvalidLmdbOpenOptions error as an Internal error - Update filter error check in tests - Introduce Terms Matching Policies fixes #2479 fixes #2484 fixes #2486 fixes #2516 fixes #2578 fixes #2580 fixes #2583 fixes #2600 fixes #2640 fixes #2672 fixes #2679 fixes #2686 # Terms Matching Policies This PR allows end users to customize matching term policies ## Todo - [x] Update the API to return the number of pages and allow users to directly choose a page instead of computing an offset - [x] Change generation of the query tree depending on the chosen settings https://github.com/meilisearch/milli/pull/598 ## Small Documentation ### Default search query request: ```sh curl \ -X POST 'http://localhost:7700/indexes/movies/search' \ -H 'Content-Type: application/json' \ --data-binary '{ "q": "doctor of tokio" }' ``` result: ```json { "hits":[...], "estimatedTotalHits":32, "query":"doctor of tokio", "limit":20, "offset":0, "processingTimeMs":7 } ``` The default behavior doesn't change with the current Meilisearch behavior: If we don't have enough documents to fit the requested limit, we remove the query words from the last to the first typed word. ## Search query with `optionalWords` parameter request: ```sh curl \ -X POST 'http://localhost:7700/indexes/movies/search' \ -H 'Content-Type: application/json' \ --data-binary '{ "q": "doctor of tokio", "matchingStrategy": "all"}' ``` result: ```json { "hits":[...], "estimatedTotalHits":1, "query":"doctor of tokio", "limit":20, "offset":0, "processingTimeMs":7 } ``` ### allowed `matchingStrategy` values #### `last` The default behavior, If we don't have enough documents to fit the requested limit, we remove the query words from the last to the first typed word. #### `all` No word will be removed, If we don't have enough documents to fit the requested limit, we return the number of documents we found. ### In charge of the feature Core: `@ManyTheFish` & `@curquiza` Docs: TBD Integration: `@bidoubiwa` Co-authored-by: ManyTheFish <many@meilisearch.com> Co-authored-by: Many the fish <many@meilisearch.com>	2022-08-23 16:21:00 +00:00
Many the fish	651a22b1ed	Enhance enum documentation Co-authored-by: Clément Renault <clement@meilisearch.com>	2022-08-23 18:11:20 +02:00
Guillaume Mourier	ff59ae56f4	cargo fmt	2022-08-23 17:17:02 +02:00
Guillaume Mourier	b2577aac52	Add suffix describing the unit when needed; Replace MeiliSearch by Meilisearch; Precised some metrics name	2022-08-23 17:09:27 +02:00
ManyTheFish	c9bb111ef3	Implement all and last matching strategy	2022-08-23 17:07:43 +02:00
ManyTheFish	e2af8dccb8	Fix filter tests	2022-08-23 16:39:39 +02:00
ManyTheFish	5e206ee84b	Classify InvalidLmdbOpenOptions as an Internal error	2022-08-23 16:39:39 +02:00
ManyTheFish	aff4b64265	Update dependencies	2022-08-23 16:39:39 +02:00
bors[bot]	0b55e7ce6a	Merge #615 615: Remove the artifacts of the past r=Kerollmops a=irevoire Co-authored-by: Irevoire <tamo@meilisearch.com>	2022-08-23 14:22:43 +00:00
Irevoire	f6024b3269	Remove the artifacts of the past	2022-08-23 16:10:38 +02:00
bors[bot]	0a2ef0037f	Merge #2689 #2690 2689: Use mimalloc as the global allocator r=Kerollmops a=loiclec milli has switched its global allocator to mimalloc already, and we have seen some performance gains as a result. Furthermore, we can use mimalloc as the global allocator on all platforms whereas jemalloc was only activated on Linux. This PR brings mimalloc to Meilisearch as well. 2690: Add LTO and codegen-units=1 to release compile options r=Kerollmops a=loiclec This PR brings Meilisearch's release compile options in line with milli (see https://github.com/meilisearch/milli/pull/606 ). Adding LTO and codegen=units=1 will make compile times longer, but they also speed up the final binary significantly. Co-authored-by: Loïc Lecrenier <loic@meilisearch.com>	2022-08-23 12:05:02 +00:00
bors[bot]	40ae26478a	Merge #2691 2691: Update version for next release (v0.29.0) in Cargo.toml files r=Kerollmops a=curquiza Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-08-23 11:41:22 +00:00
Clémentine Urquizar	6fe3f285ce	Update version for next release (v0.29.0)	2022-08-23 13:39:56 +02:00
Loïc Lecrenier	72f8adaa70	Add LTO and codegen-units=1 to release compile options	2022-08-23 13:03:57 +02:00
Loïc Lecrenier	e659c08ac4	Use mimalloc as the global allocator	2022-08-23 12:58:10 +02:00
bors[bot]	a79ff8a1a9	Merge #611 611: Upgrade charabia v0.6.0 r=curquiza a=ManyTheFish # Pull Request ## What does this PR do? - Update `log` - Upgrade `charabia` related to https://github.com/meilisearch/meilisearch/issues/2686 Co-authored-by: ManyTheFish <many@meilisearch.com>	2022-08-23 10:17:29 +00:00
bors[bot]	e314423653	Merge #613 613: Update version for next release (v0.33.0) r=Kerollmops a=curquiza Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-08-23 10:01:20 +00:00
bors[bot]	d0521e493f	Merge #612 612: Remove Bors required test for Windows r=Kerollmops a=curquiza Remove the required windows test for merging due to the issue with Lindera https://github.com/meilisearch/milli/runs/7970141278?check_suite_focus=true Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-08-23 09:47:51 +00:00
Clémentine Urquizar	9ed7324995	Update version for next release (v0.33.0)	2022-08-23 11:47:48 +02:00
Clémentine Urquizar	e140227065	Remove Bors required test for Windows	2022-08-23 11:45:29 +02:00
bors[bot]	18886dc6b7	Merge #598 598: Matching query terms policy r=Kerollmops a=ManyTheFish ## Summary Implement several optional words strategy. ## Content Replace `optional_words` boolean with an enum containing several term matching strategies: ```rust pub enum TermsMatchingStrategy { // remove last word first Last, // remove first word first First, // remove more frequent word first Frequency, // remove smallest word first Size, // only one of the word is mandatory Any, // all words are mandatory All, } ``` All strategies implemented during the prototype are kept, but only `Last` and `All` will be published by Meilisearch in the `v0.29.0` release. ## Related spec: https://github.com/meilisearch/specifications/pull/173 prototype discussion: https://github.com/meilisearch/meilisearch/discussions/2639#discussioncomment-3447699 Co-authored-by: ManyTheFish <many@meilisearch.com>	2022-08-22 15:51:37 +00:00
ManyTheFish	5391e3842c	replace optional_words by term_matching_strategy	2022-08-22 17:47:19 +02:00
ManyTheFish	f9029727e0	Fix benchmarks	2022-08-22 14:55:53 +02:00
ManyTheFish	a5b9a35c50	Activate char_map for highlighting	2022-08-22 14:39:16 +02:00
ManyTheFish	ba5ca8a362	Upgrade charabia v0.6.0	2022-08-22 14:38:00 +02:00
ManyTheFish	5943e1c3b2	Update log dependency	2022-08-22 13:55:01 +02:00
bors[bot]	ea365126b4	Merge #2657 2657: prometheus and grafana dashboards implemented r=irevoire a=pavo-tusker Implemented Basic Prometheus Metrics and Grafana Dashboard using this [Prometheus Crate](https://crates.io/crates/prometheus) [#496](https://github.com/meilisearch/product/issues/496) ![Screenshot from 2022-08-04 19-59-06](https://user-images.githubusercontent.com/43550760/182880420-71ec8591-a2cb-4fd5-b1c5-911a6dcbdaf9.png) ![Screenshot from 2022-08-04 19-58-56](https://user-images.githubusercontent.com/43550760/182880433-11727814-e230-44dd-89c9-fec3baa47b11.png) ![Screenshot from 2022-08-04 19-58-40](https://user-images.githubusercontent.com/43550760/182880436-73312a68-4f20-49f0-80e9-5e344f96db6f.png) Co-authored-by: mohandasspat <mohan.s@pavo-tusker.com> Co-authored-by: Pavo-Tusker <43550760+pavo-tusker@users.noreply.github.com> Co-authored-by: Tamo <tamo@meilisearch.com>	2022-08-22 11:15:42 +00:00
bors[bot]	b46225070f	Merge #610 610: Share heed between all sub-crates r=Kerollmops a=irevoire # Pull Request ## What does this PR do? Use the reexported version of heed in the benchmarks and the fuzzer Co-authored-by: Irevoire <tamo@meilisearch.com>	2022-08-22 08:44:31 +00:00
mohandasspat	a277cc9a18	Merge branch 'metrics/prometheus-setup' of https://github.com/pavo-tusker/meilisearch into metrics/prometheus-setup	2022-08-22 13:34:37 +05:30
mohandasspat	a37c7ba1bb	clippy & cargo fixed	2022-08-22 13:34:19 +05:30
mohandasspat	ef1d6b1694	clippy & cargo fixed	2022-08-22 13:27:26 +05:30
Tamo	099abefc6d	Merge branch 'main' into metrics/prometheus-setup	2022-08-22 09:56:15 +02:00
mohandasspat	a05101af4d	clippy & fmt fixed	2022-08-22 13:21:22 +05:30
mohandasspat	109540011a	conflict fixes	2022-08-22 13:21:22 +05:30
mohandasspat	2f92169e48	clippy issue in metrics fixed	2022-08-22 13:21:22 +05:30
Pavo-Tusker	a58b00d8f1	Update meilisearch-http/src/option.rs Co-authored-by: Tamo <irevoire@protonmail.ch>	2022-08-22 13:21:22 +05:30
mohandasspat	2b8f3c26ec	Changed prometheus metrics feature as optional	2022-08-22 13:21:22 +05:30
mohandasspat	0b6ca73790	review fixes	2022-08-22 13:21:22 +05:30
Pavo-Tusker	1f1482e97c	Update meilisearch-http/src/routes/mod.rs Co-authored-by: Tamo <irevoire@protonmail.ch>	2022-08-22 13:21:22 +05:30
mohandasspat	25fecf9360	clippy & rustfmt fixed	2022-08-22 13:21:22 +05:30
mohandasspat	4bee0565e8	prometheus and grafana dashboards implemented	2022-08-22 13:21:22 +05:30
mohandasspat	d5da063666	clippy & fmt fixed	2022-08-22 10:52:09 +05:30
mohandasspat	43bb5176a9	conflict fixes	2022-08-22 10:30:07 +05:30
Irevoire	e7624abe63	share heed between all sub-crates	2022-08-19 11:23:41 +02:00
ManyTheFish	993aa1321c	Fix query tree building	2022-08-18 17:56:06 +02:00
ManyTheFish	bff9653050	Fix remove count	2022-08-18 17:36:30 +02:00
ManyTheFish	9640976c79	Rename TermMatchingPolicies	2022-08-18 17:36:08 +02:00
bors[bot]	a0734c991c	Merge #2674 2674: Add analytics on the stats routes r=ManyTheFish a=irevoire # Pull Request ## What does this PR do? Implements https://github.com/meilisearch/specifications/pull/169 ## PR checklist Please check if your PR fulfills the following requirements: - [ ] Does this PR fix an existing issue? - [ ] Have you read the contributing guidelines? - [ ] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Irevoire <tamo@meilisearch.com>	2022-08-18 14:19:56 +00:00
bors[bot]	cb29d7d124	Merge #2678 2678: Accept either an array of documents or a single document r=irevoire a=Kerollmops # Pull Request ## What does this PR do? Fixes #2671 ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Co-authored-by: Clément Renault <clement@meilisearch.com>	2022-08-18 14:00:01 +00:00
Clément Renault	e32d5ef2b3	Fix the test with an uncomprehensible user error message	2022-08-18 14:37:44 +02:00
bors[bot]	60a7221827	Merge #609 609: Retry downloading the benchmarks datasets r=Kerollmops a=irevoire Downloading the benchmarks datasets is failing [more and more](https://github.com/meilisearch/milli/pull/607#pullrequestreview-1076023074) often; thus, instead of fixing the issue, I thought we could retry multiple times. Co-authored-by: Irevoire <tamo@meilisearch.com>	2022-08-18 11:47:09 +00:00
bors[bot]	afc10acd19	Merge #596 596: Filter operators: NOT + IN[..] r=irevoire a=loiclec # Pull Request ## What does this PR do? Implements the changes described in https://github.com/meilisearch/meilisearch/issues/2580 It is based on top of #556 Co-authored-by: Loïc Lecrenier <loic@meilisearch.com>	2022-08-18 11:24:32 +00:00
Loïc Lecrenier	c7a86b56ef	Fix filter parser compilation error	2022-08-18 13:16:56 +02:00
Loïc Lecrenier	9b6602cba2	Avoid cloning FilterCondition in filter array parsing	2022-08-18 13:06:57 +02:00
Loïc Lecrenier	8a271223a9	Change a macro_rules to a function in filter parser	2022-08-18 13:03:55 +02:00
bors[bot]	ee69ede1ce	Merge #2677 2677: Hide the batch_uid field from the tasks route r=Kerollmops a=Kerollmops # Pull Request ## What does this PR do? Fixes #2676 ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Co-authored-by: Clément Renault <clement@meilisearch.com>	2022-08-18 10:01:09 +00:00
Clément Renault	9b2036ac05	Accept either an array of documents or a single document	2022-08-18 11:55:14 +02:00
Loïc Lecrenier	dd34dbaca5	Add more filter parser tests	2022-08-18 11:55:01 +02:00
Loïc Lecrenier	5d74ebd5e5	Cargo fmt	2022-08-18 11:36:38 +02:00
Clément Renault	5c543f9d94	Add a test for single document upload	2022-08-18 11:33:22 +02:00
Loïc Lecrenier	9af69c151b	Limit the maximum depth of filters This should have no impact on the user but is there to safeguard meilisearch against malicious inputs.	2022-08-18 11:31:38 +02:00
Clément Renault	0c03ed3c1e	Hide the batch_uid field from the tasks route	2022-08-18 11:15:21 +02:00
Loïc Lecrenier	c51dcad51b	Don't recompute filterable fields in evaluation of IN[] filter	2022-08-18 10:59:21 +02:00
Loïc Lecrenier	98f0da6b38	Simplify representation of nested NOT filters	2022-08-18 10:58:24 +02:00
Loïc Lecrenier	b030efdc83	Fix parsing of IN[] filter followed by whitespace + factorise its impl	2022-08-18 10:58:04 +02:00
Irevoire	84a784834e	retry downloading the benchmarks datasets	2022-08-17 19:25:05 +02:00
bors[bot]	79094bcbcf	Merge #607 607: Better threshold r=Kerollmops a=irevoire # Pull Request ## What does this PR do? Fixes #570 This PR tries to improve the threshold used to trigger the real deletion of documents. The deletion is now triggered in two cases; - 10% of the total available space is used by soft deleted documents - 90% of the total available space is used. In this context, « total available space » means the `map_size` of lmdb. And the size used by the soft deleted documents is actually an estimation. We can't determine precisely the size used by one document thus what we do is; take the total space used, divide it by the number of documents + soft deleted documents to estimate the size of one average document. Then multiply the size of one avg document by the number of soft deleted document. -------- <img width="808" alt="image" src="https://user-images.githubusercontent.com/7032172/185083075-92cf379e-8ae1-4bfc-9ca6-93b54e6ab4e9.png"> Here we can see we have a ~10GB drift in the end between the space used by the soft deleted and the real space used by the documents. Personally I don’t think that's a big issue because once the red line reach 90GB everything will be freed but now you know. If you have an idea on how to improve this estimation I would love to hear it. It look like the difference is linear so maybe we could simply multiply the current estimation by two? Co-authored-by: Irevoire <tamo@meilisearch.com>	2022-08-17 16:31:04 +00:00
mohandasspat	54a0b47c2b	clippy issue in metrics fixed	2022-08-17 21:08:28 +05:30
Loïc Lecrenier	497f9817a2	Use snapshot testing for the filter parser	2022-08-17 17:35:01 +02:00
Pavo-Tusker	947fb5c956	Update meilisearch-http/src/option.rs Co-authored-by: Tamo <irevoire@protonmail.ch>	2022-08-17 20:57:07 +05:30
mohandasspat	cd18459484	Changed prometheus metrics feature as optional	2022-08-17 20:56:15 +05:30
mohandasspat	225d9936ed	review fixes	2022-08-17 20:55:29 +05:30
Pavo-Tusker	93daa4c464	Update meilisearch-http/src/routes/mod.rs Co-authored-by: Tamo <irevoire@protonmail.ch>	2022-08-17 20:55:29 +05:30
mohandasspat	d08c77706c	clippy & rustfmt fixed	2022-08-17 20:55:29 +05:30
mohandasspat	de58ccd4ba	prometheus and grafana dashboards implemented	2022-08-17 20:54:39 +05:30
Irevoire	4aae07d5f5	expose the size methods	2022-08-17 17:07:38 +02:00
Irevoire	e96b852107	bump heed	2022-08-17 17:05:50 +02:00
Loïc Lecrenier	238a7be58d	Fix filter parser handling of keywords and surrounding spaces Now the following fragments are allowed: AND(field = AND'field' = AND"field" =	2022-08-17 16:53:40 +02:00
Irevoire	62240b7e19	add analytics on the stats routes	2022-08-17 16:12:26 +02:00
Loïc Lecrenier	b09a8f1b91	Filters: add explicit error message when using a keyword as value	2022-08-17 16:07:00 +02:00
bors[bot]	087da5621a	Merge #587 587: Word prefix pair proximity docids indexation refactor r=Kerollmops a=loiclec # Pull Request ## What does this PR do? Refactor the code of `WordPrefixPairProximityDocIds` to make it much faster, fix a bug, and add a unit test. ## Why is it faster? Because we avoid using a sorter to insert the (`word1`, `prefix`, `proximity`) keys and their associated bitmaps, and thus we don't have to sort a potentially very big set of data. I have also added a couple of other optimisations: 1. reusing allocations 2. using a prefix trie instead of an array of prefixes to get all the prefixes of a word 3. inserting directly into the database instead of putting the data in an intermediary grenad when possible. Also avoid checking for pre-existing values in the database when we know for certain that they do not exist. ## What bug was fixed? When reindexing, the `new_prefix_fst_words` prefixes may look like: ``` ["ant", "axo", "bor"] ``` which we group by first letter: ``` [["ant", "axo"], ["bor"]] ``` Later in the code, if we have the word2 "axolotl", we try to find which subarray of prefixes contains its prefixes. This check is done with `word2.starts_with(subarray_prefixes[0])`, but `"axolotl".starts_with("ant")` is false, and thus we wrongly think that there are no prefixes in `new_prefix_fst_words` that are prefixes of `axolotl`. ## StrStrU8Codec I had to change the encoding of `StrStrU8Codec` to make the second string null-terminated as well. I don't think this should be a problem, but I may have missed some nuances about the impacts of this change. ## Requests when reviewing this PR I have explained what the code does in the module documentation of `word_pair_proximity_prefix_docids`. It would be nice if someone could read it and give their opinion on whether it is a clear explanation or not. I also have a couple questions regarding the code itself: - Should we clean up and factor out the `PrefixTrieNode` code to try and make broader use of it outside this module? For now, the prefixes undergo a few transformations: from FST, to array, to prefix trie. It seems like it could be simplified. - I wrote a function called `write_into_lmdb_database_without_merging`. (1) Are we okay with such a function existing? (2) Should it be in `grenad_helpers` instead? ## Benchmark Results We reduce the time it takes to index about 8% in most cases, but it varies between -3% and -20%. ``` group indexing_main_ce90fc62 indexing_word-prefix-pair-proximity-docids-refactor_cbad2023 ----- ---------------------- ------------------------------------------------------------ indexing/-geo-delete-facetedNumber-facetedGeo-searchable- 1.00 1893.0±233.03µs ? ?/sec 1.01 1921.2±260.79µs ? ?/sec indexing/-movies-delete-facetedString-facetedNumber-searchable- 1.05 9.4±3.51ms ? ?/sec 1.00 9.0±2.14ms ? ?/sec indexing/-movies-delete-facetedString-facetedNumber-searchable-nested- 1.22 18.3±11.42ms ? ?/sec 1.00 15.0±5.79ms ? ?/sec indexing/-songs-delete-facetedString-facetedNumber-searchable- 1.00 41.4±4.20ms ? ?/sec 1.28 53.0±13.97ms ? ?/sec indexing/-wiki-delete-searchable- 1.00 285.6±18.12ms ? ?/sec 1.03 293.1±16.09ms ? ?/sec indexing/Indexing geo_point 1.03 60.8±0.45s ? ?/sec 1.00 58.8±0.68s ? ?/sec indexing/Indexing movies in three batches 1.14 16.5±0.30s ? ?/sec 1.00 14.5±0.24s ? ?/sec indexing/Indexing movies with default settings 1.11 13.7±0.07s ? ?/sec 1.00 12.3±0.28s ? ?/sec indexing/Indexing nested movies with default settings 1.10 10.6±0.11s ? ?/sec 1.00 9.6±0.15s ? ?/sec indexing/Indexing nested movies without any facets 1.11 9.4±0.15s ? ?/sec 1.00 8.5±0.10s ? ?/sec indexing/Indexing songs in three batches with default settings 1.18 66.2±0.39s ? ?/sec 1.00 56.0±0.67s ? ?/sec indexing/Indexing songs with default settings 1.07 58.7±1.26s ? ?/sec 1.00 54.7±1.71s ? ?/sec indexing/Indexing songs without any facets 1.08 53.1±0.88s ? ?/sec 1.00 49.3±1.43s ? ?/sec indexing/Indexing songs without faceted numbers 1.08 57.7±1.33s ? ?/sec 1.00 53.3±0.98s ? ?/sec indexing/Indexing wiki 1.06 1051.1±21.46s ? ?/sec 1.00 989.6±24.55s ? ?/sec indexing/Indexing wiki in three batches 1.20 1184.8±8.93s ? ?/sec 1.00 989.7±7.06s ? ?/sec indexing/Reindexing geo_point 1.04 67.5±0.75s ? ?/sec 1.00 64.9±0.32s ? ?/sec indexing/Reindexing movies with default settings 1.12 13.9±0.17s ? ?/sec 1.00 12.4±0.13s ? ?/sec indexing/Reindexing songs with default settings 1.05 60.6±0.84s ? ?/sec 1.00 57.5±0.99s ? ?/sec indexing/Reindexing wiki 1.07 1725.0±17.92s ? ?/sec 1.00 1611.4±9.90s ? ?/sec ``` Co-authored-by: Loïc Lecrenier <loic@meilisearch.com>	2022-08-17 14:06:12 +00:00
bors[bot]	fb95e67a2a	Merge #608 608: Fix soft deleted documents r=ManyTheFish a=ManyTheFish When we replaced or updated some documents, the indexing was skipping the replaced documents. Related to https://github.com/meilisearch/meilisearch/issues/2672 Co-authored-by: ManyTheFish <many@meilisearch.com>	2022-08-17 13:38:10 +00:00
bors[bot]	e4a52e6e45	Merge #594 594: Fix(Search): Fix phrase search candidates computation r=Kerollmops a=ManyTheFish This bug is an old bug but was hidden by the proximity criterion, Phrase searches were always returning an empty candidates list when the proximity criterion is deactivated. Before the fix, we were trying to find any words[n] near words[n] instead of finding any words[n] near words[n+1], for example: for a phrase search '"Hello world"' we were searching for "hello" near "hello" first, instead of "hello" near "world". Co-authored-by: ManyTheFish <many@meilisearch.com>	2022-08-17 13:22:52 +00:00
ManyTheFish	8c3f1a9c39	Remove useless lifetime declaration	2022-08-17 15:20:43 +02:00
ManyTheFish	e9e2349ce6	Fix typo in comment	2022-08-17 15:09:48 +02:00
ManyTheFish	2668f841d1	Fix update indexing	2022-08-17 15:03:37 +02:00
ManyTheFish	7384650d85	Update test to showcase the bug	2022-08-17 15:03:08 +02:00
bors[bot]	39869be23b	Merge #590 590: Optimise facets indexing r=Kerollmops a=loiclec # Pull Request ## What does this PR do? Fixes #589 ## Notes I added documentation for the whole module which attempts to explain the shape of the databases and their purpose. However, I realise there is already some documentation about this, so I am not sure if we want to keep it. ## Benchmarks We get a ~1.15x speed up on the geo_point benchmark. ``` group indexing_main_57042355 indexing_optimise-facets-indexation_5728619a ----- ---------------------- -------------------------------------------- indexing/-geo-delete-facetedNumber-facetedGeo-searchable- 1.00 1862.7±294.45µs ? ?/sec 1.58 2.9±1.32ms ? ?/sec indexing/-movies-delete-facetedString-facetedNumber-searchable- 1.11 8.9±2.44ms ? ?/sec 1.00 8.0±1.42ms ? ?/sec indexing/-movies-delete-facetedString-facetedNumber-searchable-nested- 1.00 12.8±3.32ms ? ?/sec 1.32 16.9±6.98ms ? ?/sec indexing/-songs-delete-facetedString-facetedNumber-searchable- 1.09 43.8±4.78ms ? ?/sec 1.00 40.3±3.79ms ? ?/sec indexing/-wiki-delete-searchable- 1.08 287.4±28.72ms ? ?/sec 1.00 264.9±9.46ms ? ?/sec indexing/Indexing geo_point 1.14 61.2±0.39s ? ?/sec 1.00 53.8±0.57s ? ?/sec indexing/Indexing movies in three batches 1.00 16.6±0.12s ? ?/sec 1.00 16.5±0.10s ? ?/sec indexing/Indexing movies with default settings 1.00 14.1±0.30s ? ?/sec 1.00 14.0±0.28s ? ?/sec indexing/Indexing nested movies with default settings 1.10 10.9±0.50s ? ?/sec 1.00 10.0±0.10s ? ?/sec indexing/Indexing nested movies without any facets 1.01 9.6±0.23s ? ?/sec 1.00 9.5±0.06s ? ?/sec indexing/Indexing songs in three batches with default settings 1.07 66.3±0.55s ? ?/sec 1.00 61.8±0.63s ? ?/sec indexing/Indexing songs with default settings 1.03 58.8±0.82s ? ?/sec 1.00 57.1±1.22s ? ?/sec indexing/Indexing songs without any facets 1.00 53.6±1.09s ? ?/sec 1.01 54.0±0.58s ? ?/sec indexing/Indexing songs without faceted numbers 1.02 58.0±1.29s ? ?/sec 1.00 57.1±1.43s ? ?/sec indexing/Indexing wiki 1.00 1064.1±21.20s ? ?/sec 1.00 1068.0±20.49s ? ?/sec indexing/Indexing wiki in three batches 1.00 1182.5±9.62s ? ?/sec 1.01 1191.2±10.96s ? ?/sec indexing/Reindexing geo_point 1.12 68.0±0.21s ? ?/sec 1.00 60.5±0.82s ? ?/sec indexing/Reindexing movies with default settings 1.01 14.1±0.21s ? ?/sec 1.00 14.0±0.26s ? ?/sec indexing/Reindexing songs with default settings 1.04 61.6±0.57s ? ?/sec 1.00 59.2±0.87s ? ?/sec indexing/Reindexing wiki 1.00 1734.0±11.38s ? ?/sec 1.01 1746.6±22.48s ? ?/sec ``` Co-authored-by: Loïc Lecrenier <loic@meilisearch.com>	2022-08-17 11:46:55 +00:00
Loïc Lecrenier	6cc975704d	Add some documentation to facets.rs	2022-08-17 12:59:52 +02:00
Loïc Lecrenier	93252769af	Apply review suggestions	2022-08-17 12:41:22 +02:00
Loïc Lecrenier	196f79115a	Run cargo fmt	2022-08-17 12:28:33 +02:00
Loïc Lecrenier	d10d78d520	Add integration tests for the IN filter	2022-08-17 12:28:33 +02:00
Loïc Lecrenier	4ecfb95d0c	Improve syntax errors for `IN` filter	2022-08-17 12:28:33 +02:00
Loïc Lecrenier	2fd20fadfc	Implement the NOT IN syntax for negated IN filter	2022-08-17 12:28:33 +02:00
Loïc Lecrenier	ca97cb0eda	Implement the IN filter operator	2022-08-17 12:28:33 +02:00
Loïc Lecrenier	90a304cb07	Fix tests after simplification of NOT filter	2022-08-17 12:28:33 +02:00
Loïc Lecrenier	cc7415bb31	Simplify FilterCondition code, made possible by the new NOT operator	2022-08-17 12:28:33 +02:00
Loïc Lecrenier	44744d9e67	Implement the simplified NOT operator	2022-08-17 12:28:33 +02:00
Loïc Lecrenier	01675771d5	Reimplement `!=` filter to select all docids not selected by `=`	2022-08-17 12:28:33 +02:00
Loïc Lecrenier	258c3dd563	Make AND+OR filters n-ary (store a vector of subfilters instead of 2) NOTE: The token_at_depth is method is a bit useless now, as the only cases where there would be a toke at depth 1000 are the cases where the parser already stack-overflowed earlier. Example: (((((... (x=1) ...)))))	2022-08-17 12:28:33 +02:00
Loïc Lecrenier	39687908f1	Add documentation and comments to facets.rs	2022-08-17 12:26:49 +02:00
Loïc Lecrenier	8d4b21a005	Switch string facet levels indexation to new algo Write the algorithm once for both numbers and strings	2022-08-17 12:26:49 +02:00
Loïc Lecrenier	cf0cd92ed4	Refactor Facets::execute to increase performance	2022-08-17 12:26:49 +02:00
bors[bot]	cd2635ccfc	Merge #602 602: Use mimalloc as the default allocator r=Kerollmops a=loiclec ## What does this PR do? Use mimalloc as the global allocator for milli's benchmarks on macOS. ## Why? On Linux, we use jemalloc, which is a very fast allocator. But on macOS, we currently use the system allocator, which is very slow. In practice, this difference in allocator speed means that it is difficult to gain insight into milli's performance by running benchmarks locally on the Mac. By using mimalloc, which is another excellent allocator, we reduce the speed difference between the two platforms. Co-authored-by: Loïc Lecrenier <loic@meilisearch.com>	2022-08-17 10:26:13 +00:00
Loïc Lecrenier	78d9f0622d	cargo fmt	2022-08-17 12:21:24 +02:00
Loïc Lecrenier	4f9edf13d7	Remove commented-out function	2022-08-17 12:21:24 +02:00
Loïc Lecrenier	405555b401	Add some documentation to PrefixTrieNode	2022-08-17 12:21:24 +02:00
Loïc Lecrenier	1bc4788e59	Remove cached Allocations struct from wpppd indexing	2022-08-17 12:18:22 +02:00
Loïc Lecrenier	ef75a77464	Fix undefined behaviour caused by reusing key from the database New full snapshot: --- source: milli/src/update/word_prefix_pair_proximity_docids.rs --- 5 a 1 [101, ] 5 a 2 [101, ] 5 am 1 [101, ] 5 b 4 [101, ] 5 be 4 [101, ] am a 3 [101, ] amazing a 1 [100, ] amazing a 2 [100, ] amazing a 3 [100, ] amazing an 1 [100, ] amazing an 2 [100, ] amazing b 2 [100, ] amazing be 2 [100, ] an a 1 [100, ] an a 2 [100, 202, ] an am 1 [100, ] an an 2 [100, ] an b 3 [100, ] an be 3 [100, ] and a 2 [100, ] and a 3 [100, ] and a 4 [100, ] and am 2 [100, ] and an 3 [100, ] and b 1 [100, ] and be 1 [100, ] at a 1 [100, 202, ] at a 2 [100, 101, ] at a 3 [100, ] at am 2 [100, 101, ] at an 1 [100, 202, ] at an 3 [100, ] at b 3 [101, ] at b 4 [100, ] at be 3 [101, ] at be 4 [100, ] beautiful a 2 [100, ] beautiful a 3 [100, ] beautiful a 4 [100, ] beautiful am 3 [100, ] beautiful an 2 [100, ] beautiful an 4 [100, ] bell a 2 [101, ] bell a 4 [101, ] bell am 4 [101, ] extraordinary a 2 [202, ] extraordinary a 3 [202, ] extraordinary an 2 [202, ] house a 3 [100, 202, ] house a 4 [100, 202, ] house am 4 [100, ] house an 3 [100, 202, ] house b 2 [100, ] house be 2 [100, ] rings a 1 [101, ] rings a 3 [101, ] rings am 3 [101, ] rings b 2 [101, ] rings be 2 [101, ] the a 3 [101, ] the b 1 [101, ] the be 1 [101, ]	2022-08-17 12:17:45 +02:00
Loïc Lecrenier	7309111433	Don't run block code in doc tests of word_pair_proximity_docids	2022-08-17 12:17:18 +02:00
Loïc Lecrenier	f6f8f543e1	Run cargo fmt	2022-08-17 12:17:18 +02:00
Loïc Lecrenier	34c991ea02	Add newlines in documentation of word_prefix_pair_proximity_docids	2022-08-17 12:17:18 +02:00
Loïc Lecrenier	06f3fd8c6d	Add more comments to WordPrefixPairProximityDocids::execute	2022-08-17 12:17:18 +02:00
Loïc Lecrenier	474500362c	Update wpppd snapshots New snapshot (yes, it's wrong as well, it will get fixed later): --- source: milli/src/update/word_prefix_pair_proximity_docids.rs --- 5 a 1 [101, ] 5 a 2 [101, ] 5 am 1 [101, ] 5 b 4 [101, ] 5 be 4 [101, ] am a 3 [101, ] amazing a 1 [100, ] amazing a 2 [100, ] amazing a 3 [100, ] amazing an 1 [100, ] amazing an 2 [100, ] amazing b 2 [100, ] amazing be 2 [100, ] an a 1 [100, ] an a 2 [100, 202, ] an am 1 [100, ] an b 3 [100, ] an be 3 [100, ] and a 2 [100, ] and a 3 [100, ] and a 4 [100, ] and b 1 [100, ] and be 1 [100, ] d\0 0 [100, 202, ] an an 2 [100, ] and am 2 [100, ] and an 3 [100, ] at a 2 [100, 101, ] at a 3 [100, ] at am 2 [100, 101, ] at an 1 [100, 202, ] at an 3 [100, ] at b 3 [101, ] at b 4 [100, ] at be 3 [101, ] at be 4 [100, ] beautiful a 2 [100, ] beautiful a 3 [100, ] beautiful a 4 [100, ] beautiful am 3 [100, ] beautiful an 2 [100, ] beautiful an 4 [100, ] bell a 2 [101, ] bell a 4 [101, ] bell am 4 [101, ] extraordinary a 2 [202, ] extraordinary a 3 [202, ] extraordinary an 2 [202, ] house a 4 [100, 202, ] house a 4 [100, ] house am 4 [100, ] house an 3 [100, 202, ] house b 2 [100, ] house be 2 [100, ] rings a 1 [101, ] rings a 3 [101, ] rings am 3 [101, ] rings b 2 [101, ] rings be 2 [101, ] the a 3 [101, ] the b 1 [101, ] the be 1 [101, ]	2022-08-17 12:17:18 +02:00
Loïc Lecrenier	ea4a96761c	Move content of readme for WordPrefixPairProximityDocids into the code	2022-08-17 12:05:37 +02:00
Loïc Lecrenier	220921628b	Simplify and document WordPrefixPairProximityDocIds::execute	2022-08-17 11:59:19 +02:00
Loïc Lecrenier	044356d221	Optimise WordPrefixPairProximityDocIds merge operation	2022-08-17 11:59:18 +02:00
Loïc Lecrenier	d350114159	Add tests for WordPrefixPairProximityDocIds	2022-08-17 11:59:15 +02:00
Loïc Lecrenier	86807ca848	Refactor word prefix pair proximity indexation further	2022-08-17 11:59:13 +02:00
Loïc Lecrenier	306593144d	Refactor word prefix pair proximity indexation	2022-08-17 11:59:00 +02:00
Loïc Lecrenier	5d59bfde8a	Sort Cargo.toml dependencies	2022-08-17 11:46:56 +02:00
bors[bot]	f55034ed54	Merge #606 606: Make binaries faster on release profile through better compile options r=Kerollmops a=loiclec Using `codegen-units = 1` and `lto = 'thin'` makes the compile time a bit longer, but also produces faster binaries. I'd like to run milli's benchmark with these options, so that we can see whether it is worth enabling on meilisearch. Co-authored-by: Loïc Lecrenier <loic@meilisearch.com>	2022-08-17 08:57:24 +00:00
Loïc Lecrenier	03e679b634	Make binaries faster on release profile through better compile options	2022-08-17 10:29:33 +02:00
Loïc Lecrenier	f20e588ec1	Make sure there is one newline at eof in cargo.toml	2022-08-17 07:44:33 +02:00
Loïc Lecrenier	20be69e1b9	Always use mimalloc as the global allocator	2022-08-16 20:09:36 +02:00
bors[bot]	22874ce300	Merge #2664 2664: 🐞 fix: Support https in print_launch_resume r=irevoire a=evpeople fix #2660 # Pull Request ## What does this PR do? Fixes #2660 <!-- Please link the issue you're trying to fix with this PR, if none then please create an issue first. --> ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: evpeople <hangcaihui@gmail.com>	2022-08-16 14:40:20 +00:00
bors[bot]	b5f91b91c3	Merge #2523 2523: Improve the tasks error reporting when processed in batches r=irevoire a=Kerollmops This fixes #2478 by changing the behavior of the task handler when there is an error in a batch of document addition or update. What changes is that when there is a user error in a task in a batch we now report this task as failed with the right error message but we continue to process the other tasks. A user error can be when a geo field is invalid, a document id is invalid, or missing. fixes #2582, #2478 Co-authored-by: Kerollmops <clement@meilisearch.com> Co-authored-by: ManyTheFish <many@meilisearch.com>	2022-08-16 14:15:30 +00:00
ManyTheFish	b6e6a08f7d	Fix CI test	2022-08-16 15:14:01 +02:00
bors[bot]	8198bb9da2	Merge #2665 2665: 📎 makes clippy happy r=Kerollmops a=irevoire Co-authored-by: Irevoire <tamo@meilisearch.com>	2022-08-16 12:01:56 +00:00
bors[bot]	293a246af8	Merge #601 601: Introduce snapshot tests r=Kerollmops a=loiclec # Pull Request ## What does this PR do? Introduce snapshot tests into milli, by using the `insta` crate. This implements the idea described by #597 See: [insta.rs](https://insta.rs) ## Design There is now a new file, `snapshot_tests.rs`, which is compiled only under `#[cfg(test)]`. It exposes the `db_snap!` macro, which is used to snapshot the content of a database. When running `cargo test`, `insta` will check that the value of the current snapshot is the same as the previous one (on the file system). If they are the same, the test passes. If they are different, the test fails and you are asked to review the new snapshot to approve or reject it. We don't want to save very large snapshots to the file system, because it will pollute the git repository and increase its size too much. Instead, we only save their `md5` hashes under the name `<snapshot_name>.hash.snap`. There is a new environment variable called `MILLI_TEST_FULL_SNAPS` which can be set to `true` in order to also save the full content of the snapshot under the name `<snapshot_name>.full.snap`. However, snapshots with the extension `.full.snap` are never saved to the git repository. ## Example ```rust // In e.g. facets.rs #[test] fn my_test() { // create an index let index = TempIndex::new(): index.add_documents(...); index.update_settings(\|settings\| ...); // then snapshot the content of one of its databases // the snapshot will be saved at the current folder under facets.rs/my_test/facet_id_string_docids.snap db_snap!(index, facet_id_string_docids); index.add_documents(...); // we can also name the snapshot to ensure there is no conflict // this snapshot will be saved at facets.rs/my_test/updated/facet_id_string_docids.snap db_snap!(index, facet_id_string, docids, "updated"); // and we can also use "inline" snapshots, which insert their content in the given string literal db_snap!(index, field_distributions, `@"");` // once the snapshot is approved, it will automatically get transformed to, e.g.: // db_snap!(index, field_distributions, `@"` // my_facet 21 // other_field 3 // "); // now let's add many documents index.add_documents(...); // because the snapshot is too big, its hash is saved instead // if the MILLI_TEST_FULL_SNAPS env variable is set to true, then the full snapshot will also be saved // at facets.rs/my_test/large/facet_id_string_docids.full.snap db_snap!(index, facet_id_string_docids, "large", `@"5348bbc46b5384455b6a900666d2a502");` } ``` Co-authored-by: Loïc Lecrenier <loic@meilisearch.com>	2022-08-16 11:57:09 +00:00
Loïc Lecrenier	dea00311b6	Add type annotations to remove compiler error	2022-08-16 09:19:30 +02:00
Irevoire	68a7d6bc61	reformat	2022-08-12 15:11:01 +02:00
Irevoire	83e20027fd	📎 makes clippy happy	2022-08-12 14:18:27 +02:00
evpeople	f21a4d61da	🌈 style(http/main.rs):	2022-08-12 16:16:23 +08:00
evpeople	12538d5a44	🐞 fix: Support https when print_lanuch_resume fix #2660	2022-08-11 21:29:18 +08:00
ManyTheFish	ae174c2cca	Fix task serialization	2022-08-11 13:35:35 +02:00
bors[bot]	e6b806e0cf	Merge #2662 2662: Fix(cli): Clamp databases max size to a multiple of system page size r=Kerollmops a=ManyTheFish fix #2659 Co-authored-by: ManyTheFish <many@meilisearch.com>	2022-08-11 08:49:24 +00:00
ManyTheFish	cf955a77db	Fix(cli): Clamp databases max size to a multiple of system page size fix #2659	2022-08-11 10:44:47 +02:00
ManyTheFish	3a48de136e	Add autobatching test	2022-08-10 17:02:29 +02:00
Loïc Lecrenier	fb2b6c0c28	Use mimalloc for benchmarks on all platforms	2022-08-10 16:56:42 +02:00
Loïc Lecrenier	6f49126223	Fix db_snap macro with inline parameter	2022-08-10 15:55:22 +02:00
Loïc Lecrenier	12920f2a4f	Fix paths of snapshot tests	2022-08-10 15:53:46 +02:00
Loïc Lecrenier	4b7fd4dfae	Update insta version	2022-08-10 15:53:46 +02:00
Loïc Lecrenier	ce560fdcb5	Add documentation for `db_snap!`	2022-08-10 15:53:46 +02:00
Loïc Lecrenier	748bb86b5b	cargo fmt	2022-08-10 15:53:46 +02:00
Loïc Lecrenier	051f24f674	Switch to snapshot tests for search/matches/mod.rs	2022-08-10 15:53:46 +02:00
Loïc Lecrenier	d2e01528a6	Switch to snapshot tests for search/criteria/typo.rs	2022-08-10 15:53:46 +02:00
Loïc Lecrenier	a9c7d82693	Switch to snapshot tests for search/criteria/attribute.rs	2022-08-10 15:53:46 +02:00
Loïc Lecrenier	4bba2f41d7	Switch to snapshot tests for query_tree.rs	2022-08-10 15:53:46 +02:00
Loïc Lecrenier	8ac24d3114	Cargo fmt + fix compiler warnings/error	2022-08-10 15:53:46 +02:00
Loïc Lecrenier	6066256689	Add snapshot tests for indexing of word_prefix_pair_proximity_docids	2022-08-10 15:53:46 +02:00
Loïc Lecrenier	3a734af159	Add snapshot tests for Facets::execute	2022-08-10 15:53:46 +02:00
Loïc Lecrenier	b9907997e4	Remove old snapshot tests code	2022-08-10 15:53:46 +02:00
Loïc Lecrenier	ef889ade5d	Refactor snapshot tests	2022-08-10 15:53:46 +02:00
Loïc Lecrenier	334098a7e0	Add index snapshot test helper function	2022-08-10 15:53:46 +02:00
Loïc Lecrenier	8f73251012	Use mimalloc for benchmarks on macOS	2022-08-10 13:30:56 +02:00
ManyTheFish	b389be48a0	Factorize phrase computation	2022-08-08 10:37:31 +02:00
bors[bot]	950d8e4c44	Merge #600 600: Simplify some unit tests r=ManyTheFish a=loiclec # Pull Request ## What does this PR do? Simplify the code that is used in unit tests to create and modify an index. Basically, the following code: ```rust let path = tempfile::tempdir().unwrap(); let mut options = EnvOpenOptions::new(); options.map_size(10 * 1024 * 1024); // 10 MB let index = Index::new(options, &path).unwrap(); let mut wtxn = index.write_txn().unwrap(); let content = documents!([ { "id": 0, "name": "kevin" }, ]); let config = IndexerConfig::default(); let indexing_config = IndexDocumentsConfig::default(); let builder = IndexDocuments::new(&mut wtxn, &index, &config, indexing_config, \|_\| ()).unwrap(); let (builder, user_error) = builder.add_documents(content).unwrap(); user_error.unwrap(); builder.execute().unwrap(); wtxn.commit.unwrap(); let mut wtxn = index.write_txn().unwrap(); let config = IndexerConfig::default(); let mut builder = Settings::new(&mut wtxn, &index, &config); builder.set_primary_key(S("docid")); builder.set_filterable_fields(hashset! { S("label") }); builder.execute(\|_\| ()).unwrap(); wtxn.commit().unwrap(); ``` becomes: ```rust let index = TempIndex::new(): index.add_documents(documents!( { "id": 0, "name": "kevin" }, )).unwrap(); index.update_settings(\|settings\| { settings.set_primary_key(S("docid")); settings.set_filterable_fields(hashset! { S("label") }); }).unwrap(); ``` Then there is a bunch of options to modify the indexing configs, the map size, to reuse a transaction, etc. For example: ```rust let mut index = TempIndex::new_with_map_size(1000 * 4096 * 10); index.index_documents_config.autogenerate_docids = true; let mut wtxn = index.write_txn().unwrap(); index.update_settings_using_wtxn(&mut wtxn, \|settings\| { settings.set_primary_key(S("docids")); }).unwrap(); wtxn.commit().unwrap(); ``` Co-authored-by: Loïc Lecrenier <loic@meilisearch.com> Co-authored-by: Loïc Lecrenier <loic.lecrenier@me.com> Co-authored-by: bors[bot] <26634292+bors[bot]@users.noreply.github.com>	2022-08-04 10:19:42 +00:00
Loïc Lecrenier	58cb1c1bda	Simplify unit tests in facet/filter.rs	2022-08-04 12:03:44 +02:00
Loïc Lecrenier	acff17fb88	Simplify indexing tests	2022-08-04 12:03:13 +02:00
bors[bot]	21284cf235	Merge #556 556: Add EXISTS filter r=loiclec a=loiclec ## What does this PR do? Fixes issue [#2484](https://github.com/meilisearch/meilisearch/issues/2484) in the meilisearch repo. It creates a `field EXISTS` filter which selects all documents containing the `field` key. For example, with the following documents: ```json [{ "id": 0, "colour": [] }, { "id": 1, "colour": ["blue", "green"] }, { "id": 2, "colour": 145238 }, { "id": 3, "colour": null }, { "id": 4, "colour": { "green": [] } }, { "id": 5, "colour": {} }, { "id": 6 }] ``` Then the filter `colour EXISTS` selects the ids `[0, 1, 2, 3, 4, 5]`. The filter `colour NOT EXISTS` selects `[6]`. ## Details There is a new database named `facet-id-exists-docids`. Its keys are field ids and its values are bitmaps of all the document ids where the corresponding field exists. To create this database, the indexing part of milli had to be adapted. The implementation there is basically copy/pasted from the code handling the `facet-id-f64-docids` database, with appropriate modifications in place. There was an issue involving the flattening of documents during (re)indexing. Previously, the following JSON: ```json { "id": 0, "colour": [], "size": {} } ``` would be flattened to: ```json { "id": 0 } ``` prior to being given to the extraction pipeline. This transformation would lose the information that is needed to populate the `facet-id-exists-docids` database. Therefore, I have also changed the implementation of the `flatten-serde-json` crate. Now, as it traverses the Json, it keeps track of which key was encountered. Then, at the end, if a previously encountered key is not present in the flattened object, it adds that key to the object with an empty array as value. For example: ```json { "id": 0, "colour": { "green": [], "blue": 1 }, "size": {} } ``` becomes ```json { "id": 0, "colour": [], "colour.green": [], "colour.blue": 1, "size": [] } ``` Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-08-04 09:46:06 +00:00
bors[bot]	50f6524ff2	Merge #579 579: Stop reindexing already indexed documents r=ManyTheFish a=irevoire ``` % ./compare.sh indexing_stop-reindexing-unchanged-documents_cb5a1669.json indexing_main_eeba1960.json group indexing_main_eeba1960 indexing_stop-reindexing-unchanged-documents_cb5a1669 ----- ---------------------- ----------------------------------------------------- indexing/-geo-delete-facetedNumber-facetedGeo-searchable- 1.03 2.0±0.22ms ? ?/sec 1.00 1955.4±336.24µs ? ?/sec indexing/-movies-delete-facetedString-facetedNumber-searchable- 1.08 11.0±2.93ms ? ?/sec 1.00 10.2±4.04ms ? ?/sec indexing/-movies-delete-facetedString-facetedNumber-searchable-nested- 1.00 15.1±3.89ms ? ?/sec 1.14 17.1±5.18ms ? ?/sec indexing/-songs-delete-facetedString-facetedNumber-searchable- 1.26 59.2±12.01ms ? ?/sec 1.00 47.1±8.52ms ? ?/sec indexing/-wiki-delete-searchable- 1.08 316.6±31.53ms ? ?/sec 1.00 293.6±17.00ms ? ?/sec indexing/Indexing geo_point 1.01 60.9±0.31s ? ?/sec 1.00 60.6±0.36s ? ?/sec indexing/Indexing movies in three batches 1.04 20.0±0.30s ? ?/sec 1.00 19.2±0.25s ? ?/sec indexing/Indexing movies with default settings 1.02 19.1±0.18s ? ?/sec 1.00 18.7±0.24s ? ?/sec indexing/Indexing nested movies with default settings 1.02 26.2±0.29s ? ?/sec 1.00 25.9±0.22s ? ?/sec indexing/Indexing nested movies without any facets 1.02 25.3±0.32s ? ?/sec 1.00 24.7±0.26s ? ?/sec indexing/Indexing songs in three batches with default settings 1.00 66.7±0.41s ? ?/sec 1.01 67.1±0.86s ? ?/sec indexing/Indexing songs with default settings 1.00 58.3±0.90s ? ?/sec 1.01 58.8±1.32s ? ?/sec indexing/Indexing songs without any facets 1.00 54.5±1.43s ? ?/sec 1.01 55.2±1.29s ? ?/sec indexing/Indexing songs without faceted numbers 1.00 57.9±1.20s ? ?/sec 1.01 58.4±0.93s ? ?/sec indexing/Indexing wiki 1.00 1052.0±10.95s ? ?/sec 1.02 1069.4±20.38s ? ?/sec indexing/Indexing wiki in three batches 1.00 1193.1±8.83s ? ?/sec 1.00 1189.5±9.40s ? ?/sec indexing/Reindexing geo_point 3.22 67.5±0.73s ? ?/sec 1.00 21.0±0.16s ? ?/sec indexing/Reindexing movies with default settings 3.75 19.4±0.28s ? ?/sec 1.00 5.2±0.05s ? ?/sec indexing/Reindexing songs with default settings 8.90 61.4±0.91s ? ?/sec 1.00 6.9±0.07s ? ?/sec indexing/Reindexing wiki 1.00 1748.2±35.68s ? ?/sec 1.00 1750.5±18.53s ? ?/sec ``` tldr: We do not lose any performance on the normal indexing benchmark, but we get between 3 and 8 times faster on the reindexing benchmarks 👍 Co-authored-by: Tamo <tamo@meilisearch.com>	2022-08-04 08:10:37 +00:00
bors[bot]	e8987cf5aa	Merge #599 599: fix: Remove whitespace trimming during document id validation r=ManyTheFish a=ManyTheFish fix #592 related to https://github.com/meilisearch/meilisearch/issues/2640 Co-authored-by: ManyTheFish <many@meilisearch.com>	2022-08-03 14:55:25 +00:00
ManyTheFish	d6f9a60a32	fix: Remove whitespace trimming during document id validation fix #592	2022-08-03 11:38:40 +02:00
Tamo	7fc35c5586	remove the useless prints	2022-08-02 10:31:22 +02:00
Tamo	f156d7dd3b	Stop reindexing already indexed documents	2022-08-02 10:31:20 +02:00
bors[bot]	dfbdc565f9	Merge #2653 2653: Bump Swatinem/rust-cache from 1.4.0 to 2.0.0 r=curquiza a=dependabot[bot] Bumps [Swatinem/rust-cache](https://github.com/Swatinem/rust-cache) from 1.4.0 to 2.0.0. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/Swatinem/rust-cache/releases">Swatinem/rust-cache's releases</a>.</em></p> <blockquote> <h2>v2.0.0</h2> <ul> <li>The action code was refactored to allow for caching multiple workspaces and different <code>target</code> directory layouts.</li> <li>The <code>working-directory</code> and <code>target-dir</code> input options were replaced by a single <code>workspaces</code> option that has the form of <code>$workspace -> $target</code>.</li> <li>Support for considering <code>env-vars</code> as part of the cache key.</li> <li>The <code>sharedKey</code> input option was renamed to <code>shared-key</code> for consistency.</li> </ul> </blockquote> </details> <details> <summary>Changelog</summary> <p><em>Sourced from <a href="https://github.com/Swatinem/rust-cache/blob/master/CHANGELOG.md">Swatinem/rust-cache's changelog</a>.</em></p> <blockquote> <h2>2.0.0</h2> <ul> <li>The action code was refactored to allow for caching multiple workspaces and different <code>target</code> directory layouts.</li> <li>The <code>working-directory</code> and <code>target-dir</code> input options were replaced by a single <code>workspaces</code> option that has the form of <code>$workspace -> $target</code>.</li> <li>Support for considering <code>env-vars</code> as part of the cache key.</li> <li>The <code>sharedKey</code> input option was renamed to <code>shared-key</code> for consistency.</li> </ul> </blockquote> </details> <details> <summary>Commits</summary> <ul> <li><a href="`6720f05bc4`"><code>6720f05</code></a> 2.0.0</li> <li><a href="`5733786579`"><code>5733786</code></a> rebuild</li> <li><a href="`622616010e`"><code>6226160</code></a> prepare v2</li> <li><a href="`0497f9301f`"><code>0497f93</code></a> improve registry cleanpu</li> <li><a href="`7b8626742a`"><code>7b86267</code></a> update registry cleaning</li> <li><a href="`911d8e9e55`"><code>911d8e9</code></a> test sparse registry</li> <li><a href="`875be5ce2d`"><code>875be5c</code></a> bump cache</li> <li><a href="`07a2ee71bc`"><code>07a2ee7</code></a> lol, dependency check was reversed</li> <li><a href="`7c190ef171`"><code>7c190ef</code></a> fix actual test code ;-)</li> <li><a href="`fffd6895b2`"><code>fffd689</code></a> add some more tests</li> <li>Additional commits viewable in <a href="https://github.com/Swatinem/rust-cache/compare/v1.4.0...v2.0.0">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=Swatinem/rust-cache&package-manager=github_actions&previous-version=1.4.0&new-version=2.0.0)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) You can trigger a rebase of this PR by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) </details> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2022-08-02 07:58:48 +00:00
dependabot[bot]	1bb05f2716	Bump Swatinem/rust-cache from 1.4.0 to 2.0.0 Bumps [Swatinem/rust-cache](https://github.com/Swatinem/rust-cache) from 1.4.0 to 2.0.0. - [Release notes](https://github.com/Swatinem/rust-cache/releases) - [Changelog](https://github.com/Swatinem/rust-cache/blob/master/CHANGELOG.md) - [Commits](https://github.com/Swatinem/rust-cache/compare/v1.4.0...v2.0.0) --- updated-dependencies: - dependency-name: Swatinem/rust-cache dependency-type: direct:production update-type: version-update:semver-major ... Signed-off-by: dependabot[bot] <support@github.com>	2022-08-01 17:07:30 +00:00
ManyTheFish	e6f03f82df	Fix clippy warnings	2022-07-28 15:56:22 +02:00
ManyTheFish	58d2aad309	Change binary option and add env var support	2022-07-28 15:13:49 +02:00
Kerollmops	e3426d5b7a	Improve the tasks error reporting	2022-07-28 15:12:54 +02:00
Kerollmops	73d4869e5e	Make the changes to plug the new DocumentsBatch system	2022-07-28 14:45:33 +02:00
Kerollmops	fe32097964	Update milli v0.32	2022-07-28 14:45:10 +02:00
Loïc Lecrenier	1fe224f2c6	Update filter-parser/fuzz/.gitignore Co-authored-by: Many the fish <many@meilisearch.com>	2022-07-21 16:12:01 +02:00
bors[bot]	dc21af46e5	Merge #2634 2634: Bring `stable` changes into `main` (v0.28.1) r=ManyTheFish a=curquiza Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com> Co-authored-by: bors[bot] <26634292+bors[bot]@users.noreply.github.com> Co-authored-by: Morgane Dubus <30866152+mdubus@users.noreply.github.com>	2022-07-21 13:22:57 +00:00
Loïc Lecrenier	07003704a8	Merge branch 'filter/field-exist'	2022-07-21 14:51:41 +02:00
bors[bot]	22aa349e31	Merge #2633 2633: Fix highlight issue by updating milli to v0.31.2 r=ManyTheFish a=curquiza Fixes https://github.com/meilisearch/meilisearch/issues/2627 Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-07-21 11:08:37 +00:00
bors[bot]	e1bc610d27	Merge #595 595: Update version for next release (v0.32.0) r=ManyTheFish a=curquiza In order to release on `main` (for v0.29.0, not v0.28.1) <img width="1014" alt="Capture d’écran 2022-07-21 à 13 20 35" src="https://user-images.githubusercontent.com/20380692/180178725-381fbdf1-c0fb-4fa9-9954-452aec5a1574.png"> Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-07-21 11:07:42 +00:00
Clémentine Urquizar	9cf6acb671	Fix highlight issue by updating milli to v0.31.2	2022-07-21 14:11:24 +04:00
Clémentine Urquizar	d5e9b7305b	Update version for next release (v0.32.0)	2022-07-21 13:20:02 +04:00
ManyTheFish	cbb3b25459	Fix(Search): Fix phrase search candidates computation This bug is an old bug but was hidden by the proximity criterion, Phrase search were always returning an empty candidates list. Before the fix, we were trying to find any words[n] near words[n] instead of finding any words[n] near words[n+1], for example: for a phrase search '"Hello world"' we were searching for "hello" near "hello" first, instead of "hello" near "world".	2022-07-21 10:04:30 +02:00
bors[bot]	941af58239	Merge #561 561: Enriched documents batch reader r=curquiza a=Kerollmops ~This PR is based on #555 and must be rebased on main after it has been merged to ease the review.~ This PR contains the work in #555 and can be merged on main as soon as reviewed and approved. - [x] Create an `EnrichedDocumentsBatchReader` that contains the external documents id. - [x] Extract the primary key name and make it accessible in the `EnrichedDocumentsBatchReader`. - [x] Use the external id from the `EnrichedDocumentsBatchReader` in the `Transform::read_documents`. - [x] Remove the `update_primary_key` from the _transform.rs_ file. - [x] Really generate the auto-generated documents ids. - [x] Insert the (auto-generated) document ids in the document while processing it in `Transform::read_documents`. Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-07-21 07:08:50 +00:00
bors[bot]	78c5826c57	Merge #2631 2631: Update mini-dashboard to v0.2.1 r=curquiza a=mdubus # Pull Request ## What does this PR do? Fixes #2629 ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Morgane Dubus <30866152+mdubus@users.noreply.github.com>	2022-07-21 06:37:57 +00:00
Morgane Dubus	7e6f3274fa	Update mini-dashboard to v0.2.1	2022-07-21 09:47:07 +04:00
Loïc Lecrenier	41a0ce07cb	Add a code comment, as suggested in PR review Co-authored-by: Many the fish <many@meilisearch.com>	2022-07-20 16:20:35 +02:00
bors[bot]	94ef326be3	Merge #2630 2630: Update version for next release (v0.28.1) r=ManyTheFish a=curquiza Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-07-20 13:42:51 +00:00
Clémentine Urquizar	d01a3ab889	Update version for next release (v0.28.1)	2022-07-20 15:46:53 +04:00
Loïc Lecrenier	1506683705	Avoid using too much memory when indexing facet-exists-docids	2022-07-19 14:42:35 +02:00
Loïc Lecrenier	d0eee5ff7a	Fix compiler error	2022-07-19 13:54:30 +02:00
bors[bot]	ad494b6f77	Merge #2625 2625: Update link to Cloud beta form r=curquiza a=davelarkan # Pull Request ## What does this PR do? Fixes #2624 ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Dave Larkan <davelarkan@gmail.com>	2022-07-19 11:09:12 +00:00
Dave Larkan	2f11686c81	Update link to Cloud beta form	2022-07-19 11:10:16 +01:00
Loïc Lecrenier	aed8c69bcb	Refactor indexation of the "facet-id-exists-docids" database The idea is to directly create a sorted and merged list of bitmaps in the form of a BTreeMap<FieldId, RoaringBitmap> instead of creating a grenad::Reader where the keys are field_id and the values are docids. Then we send that BTreeMap to the thing that handles TypedChunks, which inserts its content into the database.	2022-07-19 10:07:33 +02:00
Loïc Lecrenier	1eb1e73bb3	Add integration tests for the EXISTS filter	2022-07-19 10:07:33 +02:00
Loïc Lecrenier	4f0bd317df	Remove custom implementation of BytesEncode/Decode for the FieldId	2022-07-19 10:07:33 +02:00
Loïc Lecrenier	80b962b4f4	Run cargo fmt	2022-07-19 10:07:33 +02:00
Loïc Lecrenier	ea0642c32d	Make filter parser more strict regarding spacing around operators OR, AND, NOT, TO must now be followed by spaces	2022-07-19 10:07:33 +02:00
Loïc Lecrenier	c17d616250	Refactor index_documents_check_exists_database tests	2022-07-19 10:07:33 +02:00
Loïc Lecrenier	30bd4db0fc	Simplify indexing task for facet_exists_docids database	2022-07-19 10:07:33 +02:00
Loïc Lecrenier	392472f4bb	Apply suggestions from code review Co-authored-by: Tamo <tamo@meilisearch.com>	2022-07-19 10:07:33 +02:00
Loïc Lecrenier	bd15f5625a	Fix compiler warning	2022-07-19 10:07:33 +02:00
Loïc Lecrenier	722db7b088	Ignore target directory of filter-parser/fuzz crate	2022-07-19 10:07:33 +02:00
Loïc Lecrenier	a5c9162250	Improve parser for NOT EXISTS filter Allow multiple spaces between NOT and EXISTS	2022-07-19 10:07:33 +02:00
Loïc Lecrenier	0388b2d463	Run cargo fmt	2022-07-19 10:07:33 +02:00
Loïc Lecrenier	dc64170a69	Improve syntax of EXISTS filter, allow “value NOT EXISTS”	2022-07-19 10:07:33 +02:00
Loïc Lecrenier	72452f0cb2	Implements the EXIST filter operator	2022-07-19 10:07:33 +02:00
Loïc Lecrenier	a8641b42a7	Modify flatten_serde_json to keep dummy value for all object keys Example: ```json { "id": 0, "colour" : { "green": 1 } } ``` becomes: ```json { "id": 0, "colour" : [], "colour.green": 1 } ``` to retain the information the key "colour" exists in the original json value.	2022-07-19 10:07:33 +02:00
Loïc Lecrenier	453d593ce8	Add a database containing the docids where each field exists	2022-07-19 10:07:33 +02:00
bors[bot]	5704235521	Merge #584 584: Chores: Enhance smart-crop code comments r=curquiza a=ManyTheFish Enhance explanation around smart crop algorithms Co-authored-by: ManyTheFish <many@meilisearch.com> Co-authored-by: Many the fish <many@meilisearch.com>	2022-07-19 07:08:14 +00:00
bors[bot]	f6415b679f	Merge #588 588: Fix name of "release_date" facet in movies benchmarks r=ManyTheFish a=loiclec ## What does this PR do? The `movies.json` file in the benchmark datasets contains a filterable field called "release_date", but the indexing benchmarks wrongly called the field "released_date" instead. This PR fixes that. Co-authored-by: Loïc Lecrenier <loic@meilisearch.com>	2022-07-18 15:51:09 +00:00
Many the fish	2d79720f5d	Update milli/src/search/matches/mod.rs	2022-07-18 17:48:04 +02:00
Many the fish	8ddb4e750b	Update milli/src/search/matches/mod.rs	2022-07-18 17:47:39 +02:00
Many the fish	a277daa1f2	Update milli/src/search/matches/mod.rs	2022-07-18 17:47:13 +02:00
Many the fish	fb794c6b5e	Update milli/src/search/matches/mod.rs	2022-07-18 17:46:00 +02:00
Many the fish	1237cfc249	Update milli/src/search/matches/mod.rs	2022-07-18 17:45:37 +02:00
Many the fish	d7fd5c58cd	Update milli/src/search/matches/mod.rs	2022-07-18 17:45:06 +02:00
Loïc Lecrenier	fc9f3f31e7	Change DocumentsBatchReader to access cursor and index at same time Otherwise it is not possible to iterate over all documents while using the fields index at the same time.	2022-07-18 16:08:14 +02:00
Loïc Lecrenier	ab1571cdec	Simplify Transform::read_documents, enabled by enriched documents reader	2022-07-18 12:45:47 +02:00
Loïc Lecrenier	8270e2b768	Fix name of "release_date" facet in movies benchmarks	2022-07-18 10:34:12 +02:00
Many the fish	e261ef64d7	Update milli/src/search/matches/mod.rs Co-authored-by: Clément Renault <clement@meilisearch.com>	2022-07-18 10:18:51 +02:00
Many the fish	1da4ab5918	Update milli/src/search/matches/mod.rs Co-authored-by: Clément Renault <clement@meilisearch.com>	2022-07-18 10:18:03 +02:00
Kerollmops	448114cc1c	Fix the benchmarks with the new indexation API	2022-07-12 15:22:09 +02:00
Kerollmops	25e768f31c	Fix another issue with the nested primary key selector	2022-07-12 15:14:07 +02:00
Kerollmops	192793ee38	Add some tests to check for the nested documents ids	2022-07-12 15:14:07 +02:00
Kerollmops	a892a4a79c	Introduce a function to extend from a JSON array of objects	2022-07-12 15:14:06 +02:00
Kerollmops	dc61105554	Fix the nested document id fetching function	2022-07-12 15:14:06 +02:00
Kerollmops	2eec290424	Check the validity of the latitute and longitude numbers	2022-07-12 15:14:06 +02:00
Kerollmops	5d149d631f	Remove tests for a function that no more exists	2022-07-12 15:14:06 +02:00
Kerollmops	0bbcc7b180	Expose the `DocumentId` struct to be sure to inject the generated ids	2022-07-12 15:14:06 +02:00
Kerollmops	d1a4da9812	Generate a real UUIDv4 when ids are auto-generated	2022-07-12 15:14:06 +02:00
Kerollmops	c8ebf0de47	Rename the validate function as an enriching function	2022-07-12 15:14:06 +02:00
Kerollmops	905af2a2e9	Use the primary key and external id in the transform	2022-07-12 15:14:05 +02:00
Kerollmops	742543091e	Constify the default primary key name	2022-07-12 14:55:52 +02:00
Kerollmops	5f1bfb73ee	Extract the primary key name and make it accessible	2022-07-12 14:55:52 +02:00
Kerollmops	6a0a0ae94f	Make the Transform read from an EnrichedDocumentsBatchReader	2022-07-12 14:55:52 +02:00
Kerollmops	ea852200bb	Fix the format used for a geo deleting benchmark	2022-07-12 14:55:52 +02:00
Kerollmops	dc3f092d07	Do not leak an internal grenad Error	2022-07-12 14:55:52 +02:00
Kerollmops	8ebf5eed0d	Make the nested primary key work	2022-07-12 14:55:52 +02:00
Kerollmops	19eb3b4708	Make sur that we do not accept floats as documents ids	2022-07-12 14:55:52 +02:00
Kerollmops	2ceeb51c37	Support the auto-generated ids when validating documents	2022-07-12 14:55:51 +02:00
Kerollmops	399eec5c01	Fix the indexation tests	2022-07-12 14:55:51 +02:00
Kerollmops	fcfc4caf8c	Move the Object type in the lib.rs file and use it everywhere	2022-07-12 14:55:51 +02:00
Kerollmops	0146175fe6	Introduce the validate_documents_batch function	2022-07-12 14:55:51 +02:00
Kerollmops	cefffde9af	Improve the .gitignore of the fuzz crate	2022-07-12 14:55:51 +02:00
Kerollmops	bdc4263883	Introduce the validate_documents_batch function	2022-07-12 14:55:51 +02:00
Kerollmops	a97d4d63b9	Fix the benchmarks	2022-07-12 14:55:50 +02:00
Kerollmops	f29114f94a	Fix http-ui to fit with the new DocumentsBatchBuilder/Reader structs	2022-07-12 14:52:56 +02:00
Kerollmops	a4ceef9624	Fix the cli for the new DocumentsBatchBuilder/Reader structs	2022-07-12 14:52:56 +02:00
Kerollmops	6d0498df24	Fix the fuzz tests	2022-07-12 14:52:56 +02:00
Kerollmops	e8297ad27e	Fix the tests for the new DocumentsBatchBuilder/Reader	2022-07-12 14:52:56 +02:00
Kerollmops	419ce3966c	Rework the DocumentsBatchBuilder/Reader to use grenad	2022-07-12 14:52:55 +02:00
Kerollmops	eb63af1f10	Update grenad to 0.4.2	2022-07-12 14:52:55 +02:00
Kerollmops	048e174efb	Do not allocate when parsing CSV headers	2022-07-12 14:52:55 +02:00
bors[bot]	01a47e2db5	Merge #2598 2598: Bring `stable` into `main` (v0.28.0) r=curquiza a=curquiza Co-authored-by: Tamo <tamo@meilisearch.com> Co-authored-by: bors[bot] <26634292+bors[bot]@users.noreply.github.com> Co-authored-by: Kerollmops <clement@meilisearch.com> Co-authored-by: Janith Petangoda <22471198+janithpet@users.noreply.github.com> Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com> Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com> Co-authored-by: Irevoire <tamo@meilisearch.com> Co-authored-by: ManyTheFish <many@meilisearch.com>	2022-07-11 14:06:02 +00:00
Clémentine Urquizar	cfacc79ad7	Remove duplicate step in CI	2022-07-11 15:18:15 +02:00
Clémentine Urquizar - curqui	8e370ed9ab	Merge branch 'main' into stable	2022-07-11 14:41:15 +02:00
Clémentine Urquizar	32d6af6527	Merge remote-tracking branch 'origin/release-v0.28.0' into stable	2022-07-11 14:37:21 +02:00
ManyTheFish	5d79617a56	Chores: Enhance smart-crop code comments	2022-07-07 16:28:09 +02:00
bors[bot]	be3240d2dd	Merge #2592 2592: Chores: Add a dedicated section for Language Support in the issue template r=curquiza a=ManyTheFish This new section in put upper than the feature proposal because language support is kind of a sub-category of it, and so, in the reading order, we choose to create a feature proposal only if it is not related to Language. Co-authored-by: ManyTheFish <many@meilisearch.com>	2022-07-07 12:37:36 +00:00
ManyTheFish	2de6868858	Chores: Add a dedicated section for Language Support in the issue template This new section in put apper than feature proposal because language-support is kind of a sub-category of it, and so, in the reading order, we choose to create a feature proposal only if it is not related to Language.	2022-07-07 13:48:47 +02:00
bors[bot]	d419a91207	Merge #2591 2591: Introduce the Tasks Seen event when filtering r=Kerollmops a=Kerollmops This PR fixes #2377 by introducing the Tasks Seen analytics events. Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-07-07 11:41:01 +00:00
Kerollmops	a9fb5a4d50	Introduce the Tasks Seen event when filtering	2022-07-07 11:39:23 +02:00
bors[bot]	0353537fef	Merge #2579 2579: API keys: adds action * for actions r=irevoire a=phdavis1027 # Pull Request This is PR builds on `@janithpet's` addition to DocumentsAll; it's basically a copy-and-paste job, except I used ```iter.filter()``` to avoid the possibility for duplication that they mentioned. I'm not sure how much that matters. Also, hi! This is my first open-source contribution and my first attempt to write Rust for anyone other than myself, so any feedback whatsoever is appreciated. ## What does this PR do? Fixes #2560 ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Phillip Davis <phdavis1027@gmail.com>	2022-07-07 08:54:23 +00:00
bors[bot]	da7729e4a8	Merge #2589 2589: Update create-issue-dependencies.yml r=ManyTheFish a=curquiza Minor change to update the format in the description of the issue created: remove the useless newlines See the format without this change: https://github.com/meilisearch/meilisearch/issues/2588 Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com>	2022-07-07 08:37:33 +00:00
bors[bot]	ce90fc628a	Merge #583 583: Use BufReader to read datasets in benchmarks r=ManyTheFish a=loiclec ## What does this PR do? Ensure that the datasets used by the benchmarks are read efficiently by using a `BufReader`. ## Why? Using a `BufReader` is more representative of how `meilisearch` works. It will also make performance comparisons between different branches of `milli` more accurate. Co-authored-by: Loïc Lecrenier <loic@meilisearch.com>	2022-07-07 08:13:07 +00:00
Phillip Davis	074a6a0cce	Fix typos in HeadAuthStore::put_api_key	2022-07-06 22:24:46 -04:00
Loïc Lecrenier	aae03356cb	Use BufReader to read datasets in benchmarks	2022-07-06 18:20:15 +02:00
bors[bot]	755b1a59a2	Merge #2584 2584: Format API keys in hexa instead of base64 r=curquiza a=ManyTheFish This PR: - Changes API key generation and formatting to ease the generation of the key made by our users - updates the `uuid` crate version The API key can now be generated in bash as below: ```sh echo -n $HYPHENATED_UUID \| openssl dgst -sha256 -hmac $MASTER_KEY ``` fixes the issue raised in [product/discussion#421](https://github.com/meilisearch/product/discussions/421#discussioncomment-3079410), this should not impact anything in documentation nor integration but ease the key generation on the user sides. poke `@gmourier` Co-authored-by: ManyTheFish <many@meilisearch.com>	2022-07-06 12:49:04 +00:00
Clémentine Urquizar - curqui	bb5b18b82c	Update create-issue-dependencies.yml	2022-07-06 11:01:25 +02:00
bors[bot]	01d9560318	Merge #2587 2587: Update mini-dashboard to v0.2.0 r=curquiza a=mdubus # Pull Request ## What does this PR do? Fixes #2469 ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Morgane Dubus <30866152+mdubus@users.noreply.github.com>	2022-07-06 08:54:16 +00:00
Morgane Dubus	719879d4d2	Update mini-dashboard to v0.2.0	2022-07-06 08:12:17 +02:00
Phillip Davis	fb9b298645	Leave actions as HashSet	2022-07-05 21:52:50 -04:00
Phillip Davis	23f02f241e	Run the code formatter	2022-07-05 21:08:56 -04:00
Phillip Davis	5d80ff41a2	Clean up put_key impl	2022-07-05 21:06:58 -04:00
bors[bot]	f4989590db	Merge #2585 2585: Add CI creates issue updating dependencies r=curquiza a=VasiliySoldatkin Adds CI, which uses cron GHA to create Issue "Upgrade dependencies" every 3 months Context: [#2569](https://github.com/meilisearch/meilisearch/issues/2569) Co-authored-by: Vasiliy Soldatkin <vasiliy.soldatkin@gmail.com> Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com>	2022-07-05 19:05:43 +00:00
Clémentine Urquizar - curqui	bba5fab5e5	Update .github/workflows/create-issue-dependencies.yml	2022-07-05 21:03:08 +02:00
Clémentine Urquizar - curqui	05ffe24d64	Update .github/workflows/create-issue-dependencies.yml	2022-07-05 21:02:40 +02:00
Clémentine Urquizar - curqui	6f95ae9879	Update .github/workflows/create-issue-dependencies.yml	2022-07-05 21:02:32 +02:00
Vasiliy Soldatkin	480b881e15	Remove template and add GHA from review	2022-07-05 21:08:59 +03:00
Clémentine Urquizar - curqui	43fecbf382	Update .github/workflows/create-issue-dependencies.yml	2022-07-05 19:35:09 +02:00
Clémentine Urquizar - curqui	5588a6415a	Update .github/workflows/create-issue-dependencies.yml	2022-07-05 19:34:57 +02:00
Clémentine Urquizar - curqui	b0757e75c4	Update .github/workflows/create-issue-dependencies.yml	2022-07-05 19:34:49 +02:00
bors[bot]	106be03ba8	Merge #2539 2539: Update Docker credentials r=curquiza a=curquiza This is to avoid using the tpayet credentials and use `@meili-bot` credentials instead. The `DOCKER_USERNAME` and `DOCKER_PASSWORD` are still present as secret, I will remove them once v0.28.0 is fully merged (they are still used on `release-v0.28.0`) I tested by created a tag on the branch, it worked: the tag was pushed to docker hub by meili-bot Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-07-05 17:22:26 +00:00
bors[bot]	70c55208f1	Merge #2586 2586: Fix Clippy r=irevoire a=Kerollmops This PR fixes clippy on the `release-v0.28.0` branch. Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-07-05 15:55:00 +00:00
Kerollmops	d56bf66022	Make clippy happy	2022-07-05 17:48:44 +02:00
Vasiliy Soldatkin	2c300c72c9	Add CI creates issue updating dependencies	2022-07-05 18:28:29 +03:00
bors[bot]	ebddfdb9a3	Merge #578 578: Bump uuid to 1.1.2 r=ManyTheFish a=Kerollmops Just to [align the version with Meilisearch](https://github.com/meilisearch/meilisearch/pull/2584). Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-07-05 14:56:08 +00:00
bors[bot]	eeba196053	Merge #572 572: Add reindexing benchmarks r=Kerollmops a=irevoire With #557 coming, we should add benchmarks that measure our impact on the reindexing process. Co-authored-by: Tamo <tamo@meilisearch.com>	2022-07-05 14:43:01 +00:00
Kerollmops	1bfdcfc84f	Bump uuid to 1.1.2	2022-07-05 16:23:36 +02:00
bors[bot]	dd1e606f13	Merge #557 557: Fasten documents deletion and update r=Kerollmops a=irevoire When a document deletion occurs, instead of deleting the document we mark it as deleted in the new “soft deleted” bitmap. It is then removed from the search and all the other endpoints. I ran the benchmarks against main; ``` % ./compare.sh indexing_main_83ad1aaf.json indexing_fasten-document-deletion_abab51fb.json group indexing_fasten-document-deletion_abab51fb indexing_main_83ad1aaf ----- ------------------------------------------ ---------------------- indexing/-geo-delete-facetedNumber-facetedGeo-searchable- 1.05 2.0±0.40ms ? ?/sec 1.00 1904.9±190.00µs ? ?/sec indexing/-movies-delete-facetedString-facetedNumber-searchable- 1.00 10.3±2.64ms ? ?/sec 961.61 9.9±0.12s ? ?/sec indexing/-movies-delete-facetedString-facetedNumber-searchable-nested- 1.00 15.1±3.90ms ? ?/sec 554.63 8.4±0.12s ? ?/sec indexing/-songs-delete-facetedString-facetedNumber-searchable- 1.00 45.1±7.53ms ? ?/sec 710.15 32.0±0.10s ? ?/sec indexing/-wiki-delete-searchable- 1.00 277.8±7.97ms ? ?/sec 1946.57 540.8±3.15s ? ?/sec indexing/Indexing geo_point 1.00 12.0±0.20s ? ?/sec 1.03 12.4±0.19s ? ?/sec indexing/Indexing movies in three batches 1.00 19.3±0.30s ? ?/sec 1.01 19.4±0.16s ? ?/sec indexing/Indexing movies with default settings 1.00 18.8±0.09s ? ?/sec 1.00 18.9±0.10s ? ?/sec indexing/Indexing nested movies with default settings 1.00 25.9±0.19s ? ?/sec 1.00 25.9±0.12s ? ?/sec indexing/Indexing nested movies without any facets 1.00 24.8±0.17s ? ?/sec 1.00 24.8±0.18s ? ?/sec indexing/Indexing songs in three batches with default settings 1.00 65.9±0.96s ? ?/sec 1.03 67.8±0.82s ? ?/sec indexing/Indexing songs with default settings 1.00 58.8±1.11s ? ?/sec 1.02 59.9±2.09s ? ?/sec indexing/Indexing songs without any facets 1.00 53.4±0.72s ? ?/sec 1.01 54.2±0.88s ? ?/sec indexing/Indexing songs without faceted numbers 1.00 57.9±1.17s ? ?/sec 1.01 58.3±1.20s ? ?/sec indexing/Indexing wiki 1.00 1065.2±13.26s ? ?/sec 1.00 1065.8±12.66s ? ?/sec indexing/Indexing wiki in three batches 1.00 1182.4±6.20s ? ?/sec 1.01 1190.8±8.48s ? ?/sec ``` Most things do not change, we lost 0.1ms on the indexing of geo point (I don’t get why), and then we are between 500 and 1900 times faster when we delete documents. Co-authored-by: Tamo <tamo@meilisearch.com>	2022-07-05 14:14:38 +00:00
ManyTheFish	a146fd45b9	Format API keys in hexa instead of base64	2022-07-05 16:14:18 +02:00
Tamo	250be9fe6c	put the threshold back to 10k	2022-07-05 15:57:44 +02:00
bors[bot]	62692c171d	Merge #577 577: Fix deserialisation of NDJson documents in benchmarks r=irevoire a=loiclec Previously, the first document in the NDJson file was read over and over again. So the `geo_point` benchmark was not working properly: it only indexed one document. Co-authored-by: Loïc Lecrenier <loic@meilisearch.com>	2022-07-05 13:54:47 +00:00
Loïc Lecrenier	9bc7627e27	Fix deserialisation of NDJson documents in benchmarks	2022-07-05 15:51:06 +02:00
Tamo	b61efd09fc	Makes the internal soft deleted error a UserError	2022-07-05 15:34:45 +02:00
Tamo	eaf28b0628	Apply review suggestions Co-authored-by: Clément Renault <clement@meilisearch.com>	2022-07-05 15:30:33 +02:00
Tamo	3b309f654a	Fasten the document deletion When a document deletion occurs, instead of deleting the document we mark it as deleted in the new “soft deleted” bitmap. It is then removed from the search, and all the other endpoints.	2022-07-05 15:30:33 +02:00
Tamo	2700d8dc67	Add reindexing benchmarks	2022-07-05 14:46:46 +02:00
bors[bot]	77c837fc1b	Merge #575 575: Bump charabia r=loiclec a=irevoire This fix #573 Co-authored-by: Tamo <tamo@meilisearch.com>	2022-07-05 11:53:57 +00:00
Tamo	446439e8be	bump charabia	2022-07-05 12:19:30 +02:00
Phillip Davis	63e1fb4f96	Run the code formatter	2022-07-04 21:49:40 -04:00
Phillip Davis	be1c6f9dc4	Update tests to include .* permissions for tasks, indexes, dumps, stats, and settings	2022-07-04 21:38:54 -04:00
Phillip Davis	c251b527b0	Add iterators over * for stats, dumps, tasks, settings, and indexes; change documents impl to prevent duplication	2022-07-04 21:38:31 -04:00
Phillip Davis	1dc3724c1f	Added [...]_ALL enum members in action.rs	2022-07-04 21:38:21 -04:00
bors[bot]	c1ad56281d	Merge #2545 #2556 #2568 #2573 #2574 #2575 2545: meilisearch-lib was missing a feature in its cargo.toml r=Kerollmops a=irevoire Meilisearch-lib was missing a feature to compile on its own 2556: chore: `meilisearch-http` readability improvements r=curquiza a=ryanrussell ## What does this PR do? Readability improvements in `meilisearch-http`. Believe these are pretty straightforward; let me know if anything needs adjusted or reverted :) ## PR checklist Please check if your PR fulfills the following requirements: - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? 2568: Improve manifest for dependabot r=curquiza a=curquiza Improve dependabot manifest - `rebase-strategy: disabled` -> avoid issues with bors we had in the past with the integration team. The automatic rebase of dependabot does not fit with the rebase bors will try to do - `labels` -> better triage for when I create the release changelogs 2573: Bump docker/login-action from 1 to 2 r=curquiza a=dependabot[bot] Bumps [docker/login-action](https://github.com/docker/login-action) from 1 to 2. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/docker/login-action/releases">docker/login-action's releases</a>.</em></p> <blockquote> <h2>v2.0.0</h2> <ul> <li>Node 16 as default runtime by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://github-redirect.dependabot.com/docker/login-action/issues/161">#161</a>) <ul> <li>This requires a minimum <a href="https://github.com/actions/runner/releases/tag/v2.285.0">Actions Runner</a> version of v2.285.0, which is by default available in GHES 3.4 or later.</li> </ul> </li> <li>chore: update dev dependencies and workflow by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://github-redirect.dependabot.com/docker/login-action/issues/170">#170</a>)</li> <li>Bump <code>`@actions/exec</code>` from 1.1.0 to 1.1.1 (<a href="https://github-redirect.dependabot.com/docker/login-action/issues/167">#167</a>)</li> <li>Bump <code>`@actions/io</code>` from 1.1.1 to 1.1.2 (<a href="https://github-redirect.dependabot.com/docker/login-action/issues/168">#168</a>)</li> <li>Bump minimist from 1.2.5 to 1.2.6 (<a href="https://github-redirect.dependabot.com/docker/login-action/issues/176">#176</a>)</li> <li>Bump https-proxy-agent from 5.0.0 to 5.0.1 (<a href="https://github-redirect.dependabot.com/docker/login-action/issues/182">#182</a>)</li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/login-action/compare/v1.14.1...v2.0.0">https://github.com/docker/login-action/compare/v1.14.1...v2.0.0</a></p> <h2>v1.14.1</h2> <ul> <li>Revert to Node 12 as default runtime to fix issue for GHE users (<a href="https://github-redirect.dependabot.com/docker/login-action/issues/160">#160</a>)</li> </ul> <h2>v1.14.0</h2> <ul> <li>Update to node 16 (<a href="https://github-redirect.dependabot.com/docker/login-action/issues/158">#158</a>)</li> <li>Bump <code>`@aws-sdk/client-ecr</code>` from 3.45.0 to 3.53.0 (<a href="https://github-redirect.dependabot.com/docker/login-action/issues/157">#157</a>)</li> <li>Bump <code>`@aws-sdk/client-ecr-public</code>` from 3.45.0 to 3.53.0 (<a href="https://github-redirect.dependabot.com/docker/login-action/issues/156">#156</a>)</li> </ul> <h2>v1.13.0</h2> <ul> <li>Handle proxy settings for aws-sdk (<a href="https://github-redirect.dependabot.com/docker/login-action/issues/152">#152</a>)</li> <li>Workload identity based authentication docs for GCR and GAR (<a href="https://github-redirect.dependabot.com/docker/login-action/issues/112">#112</a>)</li> <li>Test login against ACR (<a href="https://github-redirect.dependabot.com/docker/login-action/issues/49">#49</a>)</li> <li>Bump <code>`@aws-sdk/client-ecr</code>` from 3.44.0 to 3.45.0 (<a href="https://github-redirect.dependabot.com/docker/login-action/issues/132">#132</a>)</li> <li>Bump <code>`@aws-sdk/client-ecr-public</code>` from 3.43.0 to 3.45.0 (<a href="https://github-redirect.dependabot.com/docker/login-action/issues/131">#131</a>)</li> </ul> <h2>v1.12.0</h2> <ul> <li>ECR: only set credentials if username and password are specified (<a href="https://github-redirect.dependabot.com/docker/login-action/issues/128">#128</a>)</li> <li>Refactor to use aws-sdk v3 (<a href="https://github-redirect.dependabot.com/docker/login-action/issues/128">#128</a>)</li> </ul> <h2>v1.11.0</h2> <ul> <li>ECR: switch implementation to use the AWS SDK (<a href="https://github-redirect.dependabot.com/docker/login-action/issues/126">#126</a>)</li> <li><code>ecr</code> input to specify whether the given registry is ECR (<a href="https://github-redirect.dependabot.com/docker/login-action/issues/123">#123</a>)</li> <li>Test against Windows runner (<a href="https://github-redirect.dependabot.com/docker/login-action/issues/126">#126</a>)</li> <li>Update instructions for Google registry (<a href="https://github-redirect.dependabot.com/docker/login-action/issues/127">#127</a>)</li> <li>Update dev workflow (<a href="https://github-redirect.dependabot.com/docker/login-action/issues/111">#111</a>)</li> <li>Small changes for GHCR doc (<a href="https://github-redirect.dependabot.com/docker/login-action/issues/86">#86</a>)</li> <li>Update dev dependencies (<a href="https://github-redirect.dependabot.com/docker/login-action/issues/85">#85</a>)</li> <li>Bump ansi-regex from 5.0.0 to 5.0.1 (<a href="https://github-redirect.dependabot.com/docker/login-action/issues/101">#101</a>)</li> <li>Bump tmpl from 1.0.4 to 1.0.5 (<a href="https://github-redirect.dependabot.com/docker/login-action/issues/100">#100</a>)</li> <li>Bump <code>`@actions/core</code>` from 1.4.0 to 1.6.0 (<a href="https://github-redirect.dependabot.com/docker/login-action/issues/94">#94</a> <a href="https://github-redirect.dependabot.com/docker/login-action/issues/103">#103</a>)</li> <li>Bump codecov/codecov-action from 1 to 2 (<a href="https://github-redirect.dependabot.com/docker/login-action/issues/88">#88</a>)</li> <li>Bump hosted-git-info from 2.8.8 to 2.8.9 (<a href="https://github-redirect.dependabot.com/docker/login-action/issues/83">#83</a>)</li> <li>Bump node-notifier from 8.0.0 to 8.0.2 (<a href="https://github-redirect.dependabot.com/docker/login-action/issues/82">#82</a>)</li> <li>Bump ws from 7.3.1 to 7.5.0 (<a href="https://github-redirect.dependabot.com/docker/login-action/issues/81">#81</a>)</li> <li>Bump lodash from 4.17.20 to 4.17.21 (<a href="https://github-redirect.dependabot.com/docker/login-action/issues/80">#80</a>)</li> <li>Bump y18n from 4.0.0 to 4.0.3 (<a href="https://github-redirect.dependabot.com/docker/login-action/issues/79">#79</a>)</li> </ul> <h2>v1.10.0</h2> <ul> <li>GitHub Packages Docker Registry deprecated (<a href="https://github-redirect.dependabot.com/docker/login-action/issues/78">#78</a>)</li> </ul> <!-- raw HTML omitted --> </blockquote> <p>... (truncated)</p> </details> <details> <summary>Commits</summary> <ul> <li><a href="`49ed152c8e`"><code>49ed152</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/docker/login-action/issues/161">#161</a> from crazy-max/node16-runtime</li> <li><a href="`b61a9ce7bd`"><code>b61a9ce</code></a> Node 16 as default runtime</li> <li><a href="`3a136a8631`"><code>3a136a8</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/docker/login-action/issues/182">#182</a> from docker/dependabot/npm_and_yarn/https-proxy-agent...</li> <li><a href="`b312880b69`"><code>b312880</code></a> Update generated content</li> <li><a href="`795794e081`"><code>795794e</code></a> Bump https-proxy-agent from 5.0.0 to 5.0.1</li> <li><a href="`1edf6180e0`"><code>1edf618</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/docker/login-action/issues/179">#179</a> from docker/dependabot/github_actions/codecov/codecov...</li> <li><a href="`8e66ad4089`"><code>8e66ad4</code></a> Bump codecov/codecov-action from 2 to 3</li> <li><a href="`7c79b598ea`"><code>7c79b59</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/docker/login-action/issues/176">#176</a> from docker/dependabot/npm_and_yarn/minimist-1.2.6</li> <li><a href="`24a38e0d6d`"><code>24a38e0</code></a> Bump minimist from 1.2.5 to 1.2.6</li> <li><a href="`70e1ff84cb`"><code>70e1ff8</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/docker/login-action/issues/170">#170</a> from crazy-max/eslint</li> <li>Additional commits viewable in <a href="https://github.com/docker/login-action/compare/v1...v2">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=docker/login-action&package-manager=github_actions&previous-version=1&new-version=2)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) Dependabot will resolve any conflicts with this PR as long as you don't alter it yourself. You can also trigger a rebase manually by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) </details> 2574: Bump Swatinem/rust-cache from 1.3.0 to 1.4.0 r=curquiza a=dependabot[bot] Bumps [Swatinem/rust-cache](https://github.com/Swatinem/rust-cache) from 1.3.0 to 1.4.0. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/Swatinem/rust-cache/releases">Swatinem/rust-cache's releases</a>.</em></p> <blockquote> <h2>v1.4.0</h2> <ul> <li>Clean both debug and release target directories.</li> </ul> </blockquote> </details> <details> <summary>Changelog</summary> <p><em>Sourced from <a href="https://github.com/Swatinem/rust-cache/blob/v1/CHANGELOG.md">Swatinem/rust-cache's changelog</a>.</em></p> <blockquote> <h2>1.4.0</h2> <ul> <li>Clean both <code>debug</code> and <code>release</code> target directories.</li> </ul> </blockquote> </details> <details> <summary>Commits</summary> <ul> <li><a href="`cb2cf0cc7c`"><code>cb2cf0c</code></a> 1.4.0</li> <li><a href="`74e8e24b6d`"><code>74e8e24</code></a> Update dependencies, clean both debug and release targets</li> <li><a href="`f8f67b7515`"><code>f8f67b7</code></a> Add a LICENSE file</li> <li><a href="`5b2b053862`"><code>5b2b053</code></a> Improve Cache Details documentation (<a href="https://github-redirect.dependabot.com/Swatinem/rust-cache/issues/49">#49</a>)</li> <li><a href="`3bb3a9a087`"><code>3bb3a9a</code></a> update deps and rebuild</li> <li><a href="`d127014599`"><code>d127014</code></a> update dependencies</li> <li><a href="`801365cd81`"><code>801365c</code></a> hint that checkout has to be used first (<a href="https://github-redirect.dependabot.com/Swatinem/rust-cache/issues/34">#34</a>)</li> <li><a href="`c5ed9ba6b7`"><code>c5ed9ba</code></a> update dependencies and rebuild</li> <li><a href="`536c94f32c`"><code>536c94f</code></a> Cache-on-failure support (<a href="https://github-redirect.dependabot.com/Swatinem/rust-cache/issues/22">#22</a>)</li> <li>See full diff in <a href="https://github.com/Swatinem/rust-cache/compare/v1.3.0...v1.4.0">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=Swatinem/rust-cache&package-manager=github_actions&previous-version=1.3.0&new-version=1.4.0)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) Dependabot will resolve any conflicts with this PR as long as you don't alter it yourself. You can also trigger a rebase manually by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) </details> 2575: Bump docker/build-push-action from 2 to 3 r=curquiza a=dependabot[bot] Bumps [docker/build-push-action](https://github.com/docker/build-push-action) from 2 to 3. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/docker/build-push-action/releases">docker/build-push-action's releases</a>.</em></p> <blockquote> <h2>v3.0.0</h2> <ul> <li>Node 16 as default runtime by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/564">#564</a>) <ul> <li>This requires a minimum <a href="https://github.com/actions/runner/releases/tag/v2.285.0">Actions Runner</a> version of v2.285.0, which is by default available in GHES 3.4 or later.</li> </ul> </li> <li>Standalone mode support by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/601">#601</a> <a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/609">#609</a>)</li> <li>chore: update dev dependencies and workflow by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/571">#571</a>)</li> <li>Bump <code>`@actions/exec</code>` from 1.1.0 to 1.1.1 (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/573">#573</a>)</li> <li>Bump <code>`@actions/github</code>` from 5.0.0 to 5.0.1 (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/582">#582</a>)</li> <li>Bump minimist from 1.2.5 to 1.2.6 (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/584">#584</a>)</li> <li>Bump semver from 7.3.5 to 7.3.7 (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/595">#595</a>)</li> <li>Bump csv-parse from 4.16.3 to 5.0.4 (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/533">#533</a>)</li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/build-push-action/compare/v2.10.0...v3.0.0">https://github.com/docker/build-push-action/compare/v2.10.0...v3.0.0</a></p> <h2>v2.10.0</h2> <ul> <li>Add <code>imageid</code> output and use metadata to set <code>digest</code> output (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/569">#569</a>)</li> <li>Add <code>build-contexts</code> input (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/563">#563</a>)</li> <li>Enhance outputs display (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/559">#559</a>)</li> </ul> <h2>v2.9.0</h2> <ul> <li><code>add-hosts</code> input (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/553">#553</a> <a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/555">#555</a>)</li> <li>Fix git context subdir example and improve README (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/552">#552</a>)</li> <li>Add e2e tests for ACR (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/548">#548</a>)</li> <li>Add description on <code>github-token</code> option to README (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/544">#544</a>)</li> <li>Bump node-fetch from 2.6.1 to 2.6.7 (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/549">#549</a>)</li> </ul> <h2>v2.8.0</h2> <ul> <li>Allow specifying subdirectory with default git context (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/531">#531</a>)</li> <li>Add <code>cgroup-parent</code>, <code>shm-size</code>, <code>ulimit</code> inputs (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/501">#501</a>)</li> <li>Don't set outputs if empty or nil (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/470">#470</a>)</li> <li>docs: example to sanitize tags with metadata-action (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/476">#476</a>)</li> <li>docs: wrong syntax to sanitize repo slug (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/475">#475</a>)</li> <li>docs: test before pushing your image (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/455">#455</a>)</li> <li>readme: remove v1 section (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/500">#500</a>)</li> <li>ci: virtual env file system info (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/510">#510</a>)</li> <li>dev: update workflow (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/499">#499</a>)</li> <li>Bump <code>`@actions/core</code>` from 1.5.0 to 1.6.0 (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/160">#160</a>)</li> <li>Bump ansi-regex from 5.0.0 to 5.0.1 (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/469">#469</a>)</li> <li>Bump tmpl from 1.0.4 to 1.0.5 (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/465">#465</a>)</li> <li>Bump csv-parse from 4.16.0 to 4.16.3 (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/451">#451</a> <a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/459">#459</a>)</li> </ul> <h2>v2.7.0</h2> <ul> <li>Add <code>metadata</code> output (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/412">#412</a>)</li> <li>Bump <code>`@actions/core</code>` from 1.4.0 to 1.5.0 (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/439">#439</a>)</li> <li>Add note to sanitize tags (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/426">#426</a>)</li> <li>Cache backend API docs (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/406">#406</a>)</li> <li>Git context now supports subdir (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/407">#407</a>)</li> <li>Bump codecov/codecov-action from 1 to 2 (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/415">#415</a>)</li> </ul> <h2>v2.6.1</h2> <ul> <li>Small typo and ensure trimmed output (<a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/400">#400</a>)</li> </ul> <!-- raw HTML omitted --> </blockquote> <p>... (truncated)</p> </details> <details> <summary>Commits</summary> <ul> <li><a href="`e551b19e49`"><code>e551b19</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/564">#564</a> from crazy-max/node-16</li> <li><a href="`3554377aa3`"><code>3554377</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/609">#609</a> from crazy-max/ci-fix-test</li> <li><a href="`a62bc1b22b`"><code>a62bc1b</code></a> ci: fix standalone test</li> <li><a href="`c2085839e1`"><code>c208583</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/601">#601</a> from crazy-max/standalone-mode</li> <li><a href="`fcd91249e5`"><code>fcd9124</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/607">#607</a> from docker/dependabot/github_actions/docker/metadata...</li> <li><a href="`0ebe720aed`"><code>0ebe720</code></a> Bump docker/metadata-action from 3 to 4</li> <li><a href="`38b45804b5`"><code>38b4580</code></a> Standalone mode support</li> <li><a href="`ba317382dc`"><code>ba31738</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/docker/build-push-action/issues/533">#533</a> from docker/dependabot/npm_and_yarn/csv-parse-5.0.4</li> <li><a href="`43721d2346`"><code>43721d2</code></a> Update generated content</li> <li><a href="`5ea21bf2ba`"><code>5ea21bf</code></a> Fix csv-parse implementation since major update</li> <li>Additional commits viewable in <a href="https://github.com/docker/build-push-action/compare/v2...v3">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=docker/build-push-action&package-manager=github_actions&previous-version=2&new-version=3)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) Dependabot will resolve any conflicts with this PR as long as you don't alter it yourself. You can also trigger a rebase manually by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) </details> Co-authored-by: Tamo <tamo@meilisearch.com> Co-authored-by: Ryan Russell <git@ryanrussell.org> Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2022-07-04 12:13:46 +00:00
bors[bot]	83e5f45a91	Merge #2576 2576: Make clippy happy r=curquiza a=Kerollmops This PR fixes clippy. Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-07-04 11:53:27 +00:00
Kerollmops	aff8cd1774	Make clippy happy	2022-07-04 13:36:56 +02:00
dependabot[bot]	d1296d03ea	Bump docker/build-push-action from 2 to 3 Bumps [docker/build-push-action](https://github.com/docker/build-push-action) from 2 to 3. - [Release notes](https://github.com/docker/build-push-action/releases) - [Commits](https://github.com/docker/build-push-action/compare/v2...v3) --- updated-dependencies: - dependency-name: docker/build-push-action dependency-type: direct:production update-type: version-update:semver-major ... Signed-off-by: dependabot[bot] <support@github.com>	2022-07-01 17:06:44 +00:00
dependabot[bot]	6f7a4d95d9	Bump Swatinem/rust-cache from 1.3.0 to 1.4.0 Bumps [Swatinem/rust-cache](https://github.com/Swatinem/rust-cache) from 1.3.0 to 1.4.0. - [Release notes](https://github.com/Swatinem/rust-cache/releases) - [Changelog](https://github.com/Swatinem/rust-cache/blob/v1/CHANGELOG.md) - [Commits](https://github.com/Swatinem/rust-cache/compare/v1.3.0...v1.4.0) --- updated-dependencies: - dependency-name: Swatinem/rust-cache dependency-type: direct:production update-type: version-update:semver-minor ... Signed-off-by: dependabot[bot] <support@github.com>	2022-07-01 17:06:40 +00:00
dependabot[bot]	9ea96fa5c1	Bump docker/login-action from 1 to 2 Bumps [docker/login-action](https://github.com/docker/login-action) from 1 to 2. - [Release notes](https://github.com/docker/login-action/releases) - [Commits](https://github.com/docker/login-action/compare/v1...v2) --- updated-dependencies: - dependency-name: docker/login-action dependency-type: direct:production update-type: version-update:semver-major ... Signed-off-by: dependabot[bot] <support@github.com>	2022-07-01 17:06:38 +00:00
Clémentine Urquizar	19f732e623	Update manifest for dependabot	2022-06-29 11:35:41 +02:00
bors[bot]	d833e62282	Merge #2564 #2565 #2566 #2567 2564: Bump codecov/codecov-action from 1 to 3 r=curquiza a=dependabot[bot] Bumps [codecov/codecov-action](https://github.com/codecov/codecov-action) from 1 to 3. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/codecov/codecov-action/releases">codecov/codecov-action's releases</a>.</em></p> <blockquote> <h2>v3.0.0</h2> <h3>Breaking Changes</h3> <ul> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/689">#689</a> Bump to node16 and small fixes</li> </ul> <h3>Features</h3> <ul> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/688">#688</a> Incorporate <code>gcov</code> arguments for the Codecov uploader</li> </ul> <h3>Dependencies</h3> <ul> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/548">#548</a> build(deps-dev): bump jest-junit from 12.2.0 to 13.0.0</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/603">#603</a> [Snyk] Upgrade <code>`@actions/core</code>` from 1.5.0 to 1.6.0</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/628">#628</a> build(deps): bump node-fetch from 2.6.1 to 3.1.1</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/634">#634</a> build(deps): bump node-fetch from 3.1.1 to 3.2.0</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/636">#636</a> build(deps): bump openpgp from 5.0.1 to 5.1.0</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/652">#652</a> build(deps-dev): bump <code>`@vercel/ncc</code>` from 0.30.0 to 0.33.3</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/653">#653</a> build(deps-dev): bump <code>`@types/node</code>` from 16.11.21 to 17.0.18</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/659">#659</a> build(deps-dev): bump <code>`@types/jest</code>` from 27.4.0 to 27.4.1</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/667">#667</a> build(deps): bump actions/checkout from 2 to 3</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/673">#673</a> build(deps): bump node-fetch from 3.2.0 to 3.2.3</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/683">#683</a> build(deps): bump minimist from 1.2.5 to 1.2.6</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/685">#685</a> build(deps): bump <code>`@actions/github</code>` from 5.0.0 to 5.0.1</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/681">#681</a> build(deps-dev): bump <code>`@types/node</code>` from 17.0.18 to 17.0.23</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/682">#682</a> build(deps-dev): bump typescript from 4.5.5 to 4.6.3</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/676">#676</a> build(deps): bump <code>`@actions/exec</code>` from 1.1.0 to 1.1.1</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/675">#675</a> build(deps): bump openpgp from 5.1.0 to 5.2.1</li> </ul> <h2>v2.1.0</h2> <h2>2.1.0</h2> <h3>Features</h3> <ul> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/515">#515</a> Allow specifying version of Codecov uploader</li> </ul> <h3>Dependencies</h3> <ul> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/499">#499</a> build(deps-dev): bump <code>`@vercel/ncc</code>` from 0.29.0 to 0.30.0</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/508">#508</a> build(deps): bump openpgp from 5.0.0-5 to 5.0.0</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/514">#514</a> build(deps-dev): bump <code>`@types/node</code>` from 16.6.0 to 16.9.0</li> </ul> <h2>v2.0.3</h2> <h2>2.0.3</h2> <h3>Fixes</h3> <ul> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/464">#464</a> Fix wrong link in the readme</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/485">#485</a> fix: Add override OS and linux default to platform</li> </ul> <h3>Dependencies</h3> <ul> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/447">#447</a> build(deps): bump openpgp from 5.0.0-4 to 5.0.0-5</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/458">#458</a> build(deps-dev): bump eslint from 7.31.0 to 7.32.0</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/465">#465</a> build(deps-dev): bump <code>`@typescript-eslint/eslint-plugin</code>` from 4.28.4 to 4.29.1</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/466">#466</a> build(deps-dev): bump <code>`@typescript-eslint/parser</code>` from 4.28.4 to 4.29.1</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/468">#468</a> build(deps-dev): bump <code>`@types/jest</code>` from 26.0.24 to 27.0.0</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/470">#470</a> build(deps-dev): bump <code>`@types/node</code>` from 16.4.0 to 16.6.0</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/472">#472</a> build(deps): bump path-parse from 1.0.6 to 1.0.7</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/473">#473</a> build(deps-dev): bump <code>`@types/jest</code>` from 27.0.0 to 27.0.1</li> </ul> <!-- raw HTML omitted --> </blockquote> <p>... (truncated)</p> </details> <details> <summary>Changelog</summary> <p><em>Sourced from <a href="https://github.com/codecov/codecov-action/blob/master/CHANGELOG.md">codecov/codecov-action's changelog</a>.</em></p> <blockquote> <h2>3.1.0</h2> <h3>Features</h3> <ul> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/699">#699</a> Incorporate <code>xcode</code> arguments for the Codecov uploader</li> </ul> <h3>Dependencies</h3> <ul> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/694">#694</a> build(deps-dev): bump <code>`@vercel/ncc</code>` from 0.33.3 to 0.33.4</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/696">#696</a> build(deps-dev): bump <code>`@types/node</code>` from 17.0.23 to 17.0.25</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/698">#698</a> build(deps-dev): bump jest-junit from 13.0.0 to 13.2.0</li> </ul> <h2>3.0.0</h2> <h3>Breaking Changes</h3> <ul> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/689">#689</a> Bump to node16 and small fixes</li> </ul> <h3>Features</h3> <ul> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/688">#688</a> Incorporate <code>gcov</code> arguments for the Codecov uploader</li> </ul> <h3>Dependencies</h3> <ul> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/548">#548</a> build(deps-dev): bump jest-junit from 12.2.0 to 13.0.0</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/603">#603</a> [Snyk] Upgrade <code>`@actions/core</code>` from 1.5.0 to 1.6.0</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/628">#628</a> build(deps): bump node-fetch from 2.6.1 to 3.1.1</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/634">#634</a> build(deps): bump node-fetch from 3.1.1 to 3.2.0</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/636">#636</a> build(deps): bump openpgp from 5.0.1 to 5.1.0</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/652">#652</a> build(deps-dev): bump <code>`@vercel/ncc</code>` from 0.30.0 to 0.33.3</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/653">#653</a> build(deps-dev): bump <code>`@types/node</code>` from 16.11.21 to 17.0.18</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/659">#659</a> build(deps-dev): bump <code>`@types/jest</code>` from 27.4.0 to 27.4.1</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/667">#667</a> build(deps): bump actions/checkout from 2 to 3</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/673">#673</a> build(deps): bump node-fetch from 3.2.0 to 3.2.3</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/683">#683</a> build(deps): bump minimist from 1.2.5 to 1.2.6</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/685">#685</a> build(deps): bump <code>`@actions/github</code>` from 5.0.0 to 5.0.1</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/681">#681</a> build(deps-dev): bump <code>`@types/node</code>` from 17.0.18 to 17.0.23</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/682">#682</a> build(deps-dev): bump typescript from 4.5.5 to 4.6.3</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/676">#676</a> build(deps): bump <code>`@actions/exec</code>` from 1.1.0 to 1.1.1</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/675">#675</a> build(deps): bump openpgp from 5.1.0 to 5.2.1</li> </ul> <h2>2.1.0</h2> <h3>Features</h3> <ul> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/515">#515</a> Allow specifying version of Codecov uploader</li> </ul> <h3>Dependencies</h3> <ul> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/499">#499</a> build(deps-dev): bump <code>`@vercel/ncc</code>` from 0.29.0 to 0.30.0</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/508">#508</a> build(deps): bump openpgp from 5.0.0-5 to 5.0.0</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/514">#514</a> build(deps-dev): bump <code>`@types/node</code>` from 16.6.0 to 16.9.0</li> </ul> <h2>2.0.3</h2> <h3>Fixes</h3> <ul> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/464">#464</a> Fix wrong link in the readme</li> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/485">#485</a> fix: Add override OS and linux default to platform</li> </ul> <h3>Dependencies</h3> <ul> <li><a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/447">#447</a> build(deps): bump openpgp from 5.0.0-4 to 5.0.0-5</li> </ul> <!-- raw HTML omitted --> </blockquote> <p>... (truncated)</p> </details> <details> <summary>Commits</summary> <ul> <li><a href="`81cd2dc814`"><code>81cd2dc</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/699">#699</a> from codecov/feat-xcode</li> <li><a href="`a03184e530`"><code>a03184e</code></a> feat: add xcode support</li> <li><a href="`6a6a9ae7b1`"><code>6a6a9ae</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/694">#694</a> from codecov/dependabot/npm_and_yarn/vercel/ncc-0.33.4</li> <li><a href="`92a872a5e7`"><code>92a872a</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/696">#696</a> from codecov/dependabot/npm_and_yarn/types/node-17.0.25</li> <li><a href="`43a9c182dd`"><code>43a9c18</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/698">#698</a> from codecov/dependabot/npm_and_yarn/jest-junit-13.2.0</li> <li><a href="`13ce822ccd`"><code>13ce822</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/codecov/codecov-action/issues/690">#690</a> from codecov/ci-v3</li> <li><a href="`4d6dbaaea6`"><code>4d6dbaa</code></a> build(deps-dev): bump jest-junit from 13.0.0 to 13.2.0</li> <li><a href="`98f0f19300`"><code>98f0f19</code></a> build(deps-dev): bump <code>`@types/node</code>` from 17.0.23 to 17.0.25</li> <li><a href="`d3021d9910`"><code>d3021d9</code></a> build(deps-dev): bump <code>`@vercel/ncc</code>` from 0.33.3 to 0.33.4</li> <li><a href="`2c83f35c20`"><code>2c83f35</code></a> Update makefile to v3</li> <li>Additional commits viewable in <a href="https://github.com/codecov/codecov-action/compare/v1...v3">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=codecov/codecov-action&package-manager=github_actions&previous-version=1&new-version=3)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) Dependabot will resolve any conflicts with this PR as long as you don't alter it yourself. You can also trigger a rebase manually by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) </details> 2565: Bump docker/setup-qemu-action from 1 to 2 r=curquiza a=dependabot[bot] Bumps [docker/setup-qemu-action](https://github.com/docker/setup-qemu-action) from 1 to 2. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/docker/setup-qemu-action/releases">docker/setup-qemu-action's releases</a>.</em></p> <blockquote> <h2>v2.0.0</h2> <ul> <li>Node 16 as default runtime by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://github-redirect.dependabot.com/docker/setup-qemu-action/issues/48">#48</a>) <ul> <li>This requires a minimum <a href="https://github.com/actions/runner/releases/tag/v2.285.0">Actions Runner</a> version of v2.285.0, which is by default available in GHES 3.4 or later.</li> </ul> </li> <li>chore: update dev dependencies and workflow by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://github-redirect.dependabot.com/docker/setup-qemu-action/issues/43">#43</a> <a href="https://github-redirect.dependabot.com/docker/setup-qemu-action/issues/47">#47</a>)</li> <li>Bump <code>`@actions/core</code>` from 1.3.0 to 1.6.0 (<a href="https://github-redirect.dependabot.com/docker/setup-qemu-action/issues/37">#37</a> <a href="https://github-redirect.dependabot.com/docker/setup-qemu-action/issues/39">#39</a> <a href="https://github-redirect.dependabot.com/docker/setup-qemu-action/issues/41">#41</a>)</li> <li>Bump <code>`@actions/exec</code>` from 1.0.4 to 1.1.1 (<a href="https://github-redirect.dependabot.com/docker/setup-qemu-action/issues/38">#38</a> <a href="https://github-redirect.dependabot.com/docker/setup-qemu-action/issues/46">#46</a>)</li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/setup-qemu-action/compare/v1.2.0...v2.0.0">https://github.com/docker/setup-qemu-action/compare/v1.2.0...v2.0.0</a></p> <h2>v1.2.0</h2> <ul> <li>Display image information (<a href="https://github-redirect.dependabot.com/docker/setup-qemu-action/issues/36">#36</a>)</li> <li>Bump <code>`@actions/core</code>` from 1.2.7 to 1.3.0 (<a href="https://github-redirect.dependabot.com/docker/setup-qemu-action/issues/35">#35</a>)</li> </ul> <h2>v1.1.0</h2> <ul> <li>Remove os limitation (<a href="https://github-redirect.dependabot.com/docker/setup-qemu-action/issues/30">#30</a>)</li> <li>Bump <code>`@actions/core</code>` from 1.2.6 to 1.2.7 (<a href="https://github-redirect.dependabot.com/docker/setup-qemu-action/issues/29">#29</a>)</li> </ul> <h2>v1.0.2</h2> <ul> <li>Enhance workflow (<a href="https://github-redirect.dependabot.com/docker/setup-qemu-action/issues/26">#26</a>)</li> <li>Container based developer flow (<a href="https://github-redirect.dependabot.com/docker/setup-qemu-action/issues/19">#19</a> <a href="https://github-redirect.dependabot.com/docker/setup-qemu-action/issues/20">#20</a>)</li> </ul> <h2>v1.0.1</h2> <ul> <li>Fix CVE-2020-15228</li> </ul> </blockquote> </details> <details> <summary>Commits</summary> <ul> <li><a href="`8b122486ce`"><code>8b12248</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/docker/setup-qemu-action/issues/48">#48</a> from crazy-max/node-16</li> <li><a href="`466d53193c`"><code>466d531</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/docker/setup-qemu-action/issues/50">#50</a> from crazy-max/update-readme</li> <li><a href="`607c1922b5`"><code>607c192</code></a> simplify usage example</li> <li><a href="`d7849ecb9c`"><code>d7849ec</code></a> Node 16 as default runtime</li> <li><a href="`2d4bfe71c9`"><code>2d4bfe7</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/docker/setup-qemu-action/issues/47">#47</a> from crazy-max/update-dev</li> <li><a href="`224b802eb3`"><code>224b802</code></a> chore: update dev dependencies and workflow</li> <li><a href="`95bd865778`"><code>95bd865</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/docker/setup-qemu-action/issues/46">#46</a> from docker/dependabot/npm_and_yarn/actions/exec-1.1.1</li> <li><a href="`cfd091faa1`"><code>cfd091f</code></a> Bump <code>`@actions/exec</code>` from 1.1.0 to 1.1.1</li> <li><a href="`d2a60302b8`"><code>d2a6030</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/docker/setup-qemu-action/issues/45">#45</a> from docker/dependabot/github_actions/actions/checkout-3</li> <li><a href="`97dc484a91`"><code>97dc484</code></a> Bump actions/checkout from 2 to 3</li> <li>Additional commits viewable in <a href="https://github.com/docker/setup-qemu-action/compare/v1...v2">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=docker/setup-qemu-action&package-manager=github_actions&previous-version=1&new-version=2)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) Dependabot will resolve any conflicts with this PR as long as you don't alter it yourself. You can also trigger a rebase manually by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) </details> 2566: Bump actions/checkout from 2 to 3 r=curquiza a=dependabot[bot] Bumps [actions/checkout](https://github.com/actions/checkout) from 2 to 3. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/actions/checkout/releases">actions/checkout's releases</a>.</em></p> <blockquote> <h2>v3.0.0</h2> <ul> <li>Updated to the node16 runtime by default <ul> <li>This requires a minimum <a href="https://github.com/actions/runner/releases/tag/v2.285.0">Actions Runner</a> version of v2.285.0 to run, which is by default available in GHES 3.4 or later.</li> </ul> </li> </ul> <h2>v2.4.2</h2> <h2>What's Changed</h2> <ul> <li>Add set-safe-directory input to allow customers to take control. (<a href="https://github-redirect.dependabot.com/actions/checkout/issues/770">#770</a>) by <a href="https://github.com/TingluoHuang"><code>`@TingluoHuang</code></a>` in <a href="https://github-redirect.dependabot.com/actions/checkout/pull/776">actions/checkout#776</a></li> <li>Prepare changelog for v2.4.2. by <a href="https://github.com/TingluoHuang"><code>`@TingluoHuang</code></a>` in <a href="https://github-redirect.dependabot.com/actions/checkout/pull/778">actions/checkout#778</a></li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/actions/checkout/compare/v2...v2.4.2">https://github.com/actions/checkout/compare/v2...v2.4.2</a></p> <h2>v2.4.1</h2> <ul> <li>Fixed an issue where checkout failed to run in container jobs due to the new git setting <code>safe.directory</code></li> </ul> <h2>v2.4.0</h2> <ul> <li>Convert SSH URLs like <code>org-<ORG_ID>`@github.com:</code>` to <code>https://github.com/</code> - <a href="https://github-redirect.dependabot.com/actions/checkout/pull/621">pr</a></li> </ul> <h2>v2.3.5</h2> <p>Update dependencies</p> <h2>v2.3.4</h2> <ul> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/379">Add missing <code>await</code>s</a></li> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/360">Swap to Environment Files</a></li> </ul> <h2>v2.3.3</h2> <ul> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/345">Remove Unneeded commit information from build logs</a></li> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/326">Add Licensed to verify third party dependencies</a></li> </ul> <h2>v2.3.2</h2> <p><a href="https://github-redirect.dependabot.com/actions/checkout/pull/320">Add Third Party License Information to Dist Files</a></p> <h2>v2.3.1</h2> <p><a href="https://github-redirect.dependabot.com/actions/checkout/pull/284">Fix default branch resolution for .wiki and when using SSH</a></p> <h2>v2.3.0</h2> <p><a href="https://github-redirect.dependabot.com/actions/checkout/pull/278">Fallback to the default branch</a></p> <h2>v2.2.0</h2> <p><a href="https://github-redirect.dependabot.com/actions/checkout/pull/258">Fetch all history for all tags and branches when fetch-depth=0</a></p> <h2>v2.1.1</h2> <p>Changes to support GHES (<a href="https://github-redirect.dependabot.com/actions/checkout/pull/236">here</a> and <a href="https://github-redirect.dependabot.com/actions/checkout/pull/248">here</a>)</p> <h2>v2.1.0</h2> <ul> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/191">Group output</a></li> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/199">Changes to support GHES alpha release</a></li> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/184">Persist core.sshCommand for submodules</a></li> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/163">Add support ssh</a></li> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/179">Convert submodule SSH URL to HTTPS, when not using SSH</a></li> </ul> <!-- raw HTML omitted --> </blockquote> <p>... (truncated)</p> </details> <details> <summary>Changelog</summary> <p><em>Sourced from <a href="https://github.com/actions/checkout/blob/main/CHANGELOG.md">actions/checkout's changelog</a>.</em></p> <blockquote> <h1>Changelog</h1> <h2>v3.0.2</h2> <ul> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/770">Add input <code>set-safe-directory</code></a></li> </ul> <h2>v3.0.1</h2> <ul> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/762">Fixed an issue where checkout failed to run in container jobs due to the new git setting <code>safe.directory</code></a></li> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/744">Bumped various npm package versions</a></li> </ul> <h2>v3.0.0</h2> <ul> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/689">Update to node 16</a></li> </ul> <h2>v2.3.1</h2> <ul> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/284">Fix default branch resolution for .wiki and when using SSH</a></li> </ul> <h2>v2.3.0</h2> <ul> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/278">Fallback to the default branch</a></li> </ul> <h2>v2.2.0</h2> <ul> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/258">Fetch all history for all tags and branches when fetch-depth=0</a></li> </ul> <h2>v2.1.1</h2> <ul> <li>Changes to support GHES (<a href="https://github-redirect.dependabot.com/actions/checkout/pull/236">here</a> and <a href="https://github-redirect.dependabot.com/actions/checkout/pull/248">here</a>)</li> </ul> <h2>v2.1.0</h2> <ul> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/191">Group output</a></li> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/199">Changes to support GHES alpha release</a></li> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/184">Persist core.sshCommand for submodules</a></li> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/163">Add support ssh</a></li> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/179">Convert submodule SSH URL to HTTPS, when not using SSH</a></li> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/157">Add submodule support</a></li> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/144">Follow proxy settings</a></li> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/141">Fix ref for pr closed event when a pr is merged</a></li> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/128">Fix issue checking detached when git less than 2.22</a></li> </ul> <h2>v2.0.0</h2> <ul> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/108">Do not pass cred on command line</a></li> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/107">Add input persist-credentials</a></li> <li><a href="https://github-redirect.dependabot.com/actions/checkout/pull/104">Fallback to REST API to download repo</a></li> </ul> </blockquote> </details> <details> <summary>Commits</summary> <ul> <li><a href="`2541b1294d`"><code>2541b12</code></a> Prepare changelog for v3.0.2. (<a href="https://github-redirect.dependabot.com/actions/checkout/issues/777">#777</a>)</li> <li><a href="`0ffe6f9c55`"><code>0ffe6f9</code></a> Add set-safe-directory input to allow customers to take control. (<a href="https://github-redirect.dependabot.com/actions/checkout/issues/770">#770</a>)</li> <li><a href="`dcd71f6466`"><code>dcd71f6</code></a> Enforce safe directory (<a href="https://github-redirect.dependabot.com/actions/checkout/issues/762">#762</a>)</li> <li><a href="`add3486cc3`"><code>add3486</code></a> Patch to fix the dependbot alert. (<a href="https://github-redirect.dependabot.com/actions/checkout/issues/744">#744</a>)</li> <li><a href="`5126516654`"><code>5126516</code></a> Bump minimist from 1.2.5 to 1.2.6 (<a href="https://github-redirect.dependabot.com/actions/checkout/issues/741">#741</a>)</li> <li><a href="`d50f8ea767`"><code>d50f8ea</code></a> Add v3.0 release information to changelog (<a href="https://github-redirect.dependabot.com/actions/checkout/issues/740">#740</a>)</li> <li><a href="`2d1c1198e7`"><code>2d1c119</code></a> update test workflows to checkout v3 (<a href="https://github-redirect.dependabot.com/actions/checkout/issues/709">#709</a>)</li> <li><a href="`a12a3943b4`"><code>a12a394</code></a> update readme for v3 (<a href="https://github-redirect.dependabot.com/actions/checkout/issues/708">#708</a>)</li> <li><a href="`8f9e05e482`"><code>8f9e05e</code></a> Update to node 16 (<a href="https://github-redirect.dependabot.com/actions/checkout/issues/689">#689</a>)</li> <li>See full diff in <a href="https://github.com/actions/checkout/compare/v2...v3">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=actions/checkout&package-manager=github_actions&previous-version=2&new-version=3)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) Dependabot will resolve any conflicts with this PR as long as you don't alter it yourself. You can also trigger a rebase manually by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) </details> 2567: Bump docker/setup-buildx-action from 1 to 2 r=curquiza a=dependabot[bot] Bumps [docker/setup-buildx-action](https://github.com/docker/setup-buildx-action) from 1 to 2. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/docker/setup-buildx-action/releases">docker/setup-buildx-action's releases</a>.</em></p> <blockquote> <h2>v2.0.0</h2> <ul> <li>Node 16 as default runtime by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://github-redirect.dependabot.com/docker/setup-buildx-action/issues/131">#131</a>) <ul> <li>This requires a minimum <a href="https://github.com/actions/runner/releases/tag/v2.285.0">Actions Runner</a> version of v2.285.0, which is by default available in GHES 3.4 or later.</li> </ul> </li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/setup-buildx-action/compare/v1.7.0...v2.0.0">https://github.com/docker/setup-buildx-action/compare/v1.7.0...v2.0.0</a></p> <h2>v1.7.0</h2> <ul> <li>Standalone mode by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` in (<a href="https://github-redirect.dependabot.com/docker/setup-buildx-action/issues/119">#119</a>)</li> <li>Update dev dependencies and workflow by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://github-redirect.dependabot.com/docker/setup-buildx-action/issues/114">#114</a> <a href="https://github-redirect.dependabot.com/docker/setup-buildx-action/issues/130">#130</a>)</li> <li>Bump tmpl from 1.0.4 to 1.0.5 (<a href="https://github-redirect.dependabot.com/docker/setup-buildx-action/issues/108">#108</a>)</li> <li>Bump ansi-regex from 5.0.0 to 5.0.1 (<a href="https://github-redirect.dependabot.com/docker/setup-buildx-action/issues/109">#109</a>)</li> <li>Bump <code>`@actions/core</code>` from 1.5.0 to 1.6.0 (<a href="https://github-redirect.dependabot.com/docker/setup-buildx-action/issues/110">#110</a>)</li> <li>Bump actions/checkout from 2 to 3 (<a href="https://github-redirect.dependabot.com/docker/setup-buildx-action/issues/126">#126</a>)</li> <li>Bump <code>`@actions/tool-cache</code>` from 1.7.1 to 1.7.2 (<a href="https://github-redirect.dependabot.com/docker/setup-buildx-action/issues/128">#128</a>)</li> <li>Bump <code>`@actions/exec</code>` from 1.1.0 to 1.1.1 (<a href="https://github-redirect.dependabot.com/docker/setup-buildx-action/issues/129">#129</a>)</li> <li>Bump minimist from 1.2.5 to 1.2.6 (<a href="https://github-redirect.dependabot.com/docker/setup-buildx-action/issues/132">#132</a>)</li> <li>Bump codecov/codecov-action from 2 to 3 (<a href="https://github-redirect.dependabot.com/docker/setup-buildx-action/issues/133">#133</a>)</li> <li>Bump semver from 7.3.5 to 7.3.7 (<a href="https://github-redirect.dependabot.com/docker/setup-buildx-action/issues/136">#136</a>)</li> </ul> <h2>v1.6.0</h2> <ul> <li>Add <code>config-inline</code> input (<a href="https://github-redirect.dependabot.com/docker/setup-buildx-action/issues/106">#106</a>)</li> <li>Bump <code>`@actions/core</code>` from 1.4.0 to 1.5.0 (<a href="https://github-redirect.dependabot.com/docker/setup-buildx-action/issues/104">#104</a>)</li> <li>Bump codecov/codecov-action from 1 to 2 (<a href="https://github-redirect.dependabot.com/docker/setup-buildx-action/issues/101">#101</a>)</li> </ul> <h2>v1.5.1</h2> <ul> <li>Explicit version spec for caching (<a href="https://github-redirect.dependabot.com/docker/setup-buildx-action/issues/100">#100</a>)</li> </ul> <h2>v1.5.0</h2> <ul> <li>Allow building buildx from source (<a href="https://github-redirect.dependabot.com/docker/setup-buildx-action/issues/99">#99</a>)</li> </ul> <h2>v1.4.1</h2> <ul> <li>Fix <code>docker: invalid reference format</code> (<a href="https://github-redirect.dependabot.com/docker/setup-buildx-action/issues/97">#97</a>)</li> </ul> <h2>v1.4.0</h2> <ul> <li>Update dev deps (<a href="https://github-redirect.dependabot.com/docker/setup-buildx-action/issues/95">#95</a>)</li> <li>Use built-in <code>getExecOutput</code> (<a href="https://github-redirect.dependabot.com/docker/setup-buildx-action/issues/94">#94</a>)</li> <li>Use <code>core.getBooleanInput</code> (<a href="https://github-redirect.dependabot.com/docker/setup-buildx-action/issues/93">#93</a>)</li> <li>Bump <code>`@actions/exec</code>` from 1.0.4 to 1.1.0 (<a href="https://github-redirect.dependabot.com/docker/setup-buildx-action/issues/85">#85</a>)</li> <li>Bump y18n from 4.0.0 to 4.0.3 (<a href="https://github-redirect.dependabot.com/docker/setup-buildx-action/issues/91">#91</a>)</li> <li>Bump hosted-git-info from 2.8.8 to 2.8.9 (<a href="https://github-redirect.dependabot.com/docker/setup-buildx-action/issues/89">#89</a>)</li> <li>Bump ws from 7.3.1 to 7.5.0 (<a href="https://github-redirect.dependabot.com/docker/setup-buildx-action/issues/90">#90</a>)</li> <li>Bump <code>`@actions/tool-cache</code>` from 1.6.1 to 1.7.1 (<a href="https://github-redirect.dependabot.com/docker/setup-buildx-action/issues/82">#82</a> <a href="https://github-redirect.dependabot.com/docker/setup-buildx-action/issues/86">#86</a>)</li> <li>Bump <code>`@actions/core</code>` from 1.2.7 to 1.4.0 (<a href="https://github-redirect.dependabot.com/docker/setup-buildx-action/issues/80">#80</a> <a href="https://github-redirect.dependabot.com/docker/setup-buildx-action/issues/87">#87</a>)</li> </ul> <h2>v1.3.0</h2> <ul> <li>Display BuildKit version (<a href="https://github-redirect.dependabot.com/docker/setup-buildx-action/issues/72">#72</a>)</li> </ul> <h2>v1.2.0</h2> <ul> <li>Remove os limitation (<a href="https://github-redirect.dependabot.com/docker/setup-buildx-action/issues/71">#71</a>)</li> <li>Add test job for <code>config</code> input (<a href="https://github-redirect.dependabot.com/docker/setup-buildx-action/issues/68">#68</a>)</li> </ul> <!-- raw HTML omitted --> </blockquote> <p>... (truncated)</p> </details> <details> <summary>Commits</summary> <ul> <li><a href="`dc7b9719a9`"><code>dc7b971</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/docker/setup-buildx-action/issues/131">#131</a> from crazy-max/node16</li> <li><a href="`f55bc08278`"><code>f55bc08</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/docker/setup-buildx-action/issues/141">#141</a> from crazy-max/fix-test</li> <li><a href="`aa877a9d36`"><code>aa877a9</code></a> ci: fix standalone test</li> <li><a href="`130c56f342`"><code>130c56f</code></a> Node 16 as default runtime</li> <li>See full diff in <a href="https://github.com/docker/setup-buildx-action/compare/v1...v2">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=docker/setup-buildx-action&package-manager=github_actions&previous-version=1&new-version=2)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) Dependabot will resolve any conflicts with this PR as long as you don't alter it yourself. You can also trigger a rebase manually by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) </details> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2022-06-29 08:34:57 +00:00
bors[bot]	a733271ced	Merge #2563 2563: Bump docker/metadata-action from 3 to 4 r=curquiza a=dependabot[bot] Bumps [docker/metadata-action](https://github.com/docker/metadata-action) from 3 to 4. <details> <summary>Release notes</summary> <p><em>Sourced from <a href="https://github.com/docker/metadata-action/releases">docker/metadata-action's releases</a>.</em></p> <blockquote> <h2>v4.0.0</h2> <ul> <li>Node 16 as default runtime by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/176">#176</a>) <ul> <li>This requires a minimum <a href="https://github.com/actions/runner/releases/tag/v2.285.0">Actions Runner</a> version of v2.285.0, which is by default available in GHES 3.4 or later.</li> </ul> </li> <li>Do not sanitize before pattern matching by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/201">#201</a>) <ul> <li>Breaking change with <code>type=match</code> pattern matching</li> </ul> </li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/metadata-action/compare/v3.8.0...v4.0.0">https://github.com/docker/metadata-action/compare/v3.8.0...v4.0.0</a></p> <h2>v3.8.0</h2> <ul> <li>Add attribute to enable/disable images by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/193">#193</a>)</li> <li>Add <code>is_default_branch</code> global expression by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/192">#192</a> <a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/197">#197</a> <a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/198">#198</a>)</li> <li>Update fixtures (dev) by <a href="https://github.com/crazy-max"><code>`@crazy-max</code></a>` (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/190">#190</a>)</li> <li>Bump semver from 7.3.5 to 7.3.7 (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/185">#185</a>)</li> <li>Bump moment from 2.29.2 to 2.29.3 (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/187">#187</a>)</li> <li>Bump csv-parse from 4.16.3 to 5.0.4 (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/195">#195</a>)</li> </ul> <p><strong>Full Changelog</strong>: <a href="https://github.com/docker/metadata-action/compare/v3.7.0...v3.8.0">https://github.com/docker/metadata-action/compare/v3.7.0...v3.8.0</a></p> <h2>v3.7.0</h2> <ul> <li>Handle comments for multi-line inputs (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/172">#172</a>)</li> <li>Missing <code>json</code> output in <code>action.yml</code> (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/167">#167</a>)</li> <li>Update dev dependencies and workflow (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/175">#175</a>)</li> <li>Bump minimist from 1.2.5 to 1.2.6 (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/182">#182</a>)</li> <li>Bump moment from 2.29.1 to 2.29.2 (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/180">#180</a>)</li> <li>Bump <code>`@actions/github</code>` from 5.0.0 to 5.0.1 (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/179">#179</a>)</li> <li>Bump node-fetch from 2.6.1 to 2.6.7 (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/173">#173</a>)</li> </ul> <h2>v3.6.2</h2> <ul> <li>Handle raw statement for pre-release (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/155">#155</a> <a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/156">#156</a>)</li> </ul> <h2>v3.6.1</h2> <ul> <li>Preserve quotes inside unquoted field (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/153">#153</a>)</li> </ul> <h2>v3.6.0</h2> <ul> <li><code>base_ref</code> global expression (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/142">#142</a>)</li> <li>Trim tags and flavor inputs (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/143">#143</a>)</li> <li>Bump <code>`@actions/core</code>` from 1.5.0 to 1.6.0 (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/135">#135</a>)</li> <li>Bump ansi-regex from 5.0.0 to 5.0.1 (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/134">#134</a>)</li> <li>Bump tmpl from 1.0.4 to 1.0.5 (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/132">#132</a>)</li> <li>Bump csv-parse from 4.16.0 to 4.16.3 (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/131">#131</a>)</li> </ul> <h2>v3.5.0</h2> <ul> <li>Add global expression <code>date</code> (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/121">#121</a>)</li> <li>Bump <code>`@actions/core</code>` from 1.4.0 to 1.5.0 (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/122">#122</a>)</li> </ul> <h2>v3.4.1</h2> <ul> <li>Only return edge if branch matches (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/115">#115</a>)</li> </ul> <h2>v3.4.0</h2> <ul> <li>PEP 440 support (<a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/108">#108</a>)</li> </ul> <!-- raw HTML omitted --> </blockquote> <p>... (truncated)</p> </details> <details> <summary>Upgrade guide</summary> <p><em>Sourced from <a href="https://github.com/docker/metadata-action/blob/master/UPGRADE.md">docker/metadata-action's upgrade guide</a>.</em></p> <blockquote> <h1>Upgrade notes</h1> <h2>v2 to v3</h2> <ul> <li>Repository has been moved to docker org. Replace <code>crazy-max/ghaction-docker-meta@v2</code> with <code>docker/metadata-action@v4</code></li> <li>The default bake target has been changed: <code>ghaction-docker-meta</code> > <code>docker-metadata-action</code></li> </ul> <h2>v1 to v2</h2> <ul> <li><a href="https://github.com/docker/metadata-action/blob/master/#inputs">inputs</a> <ul> <li><a href="https://github.com/docker/metadata-action/blob/master/#tag-sha"><code>tag-sha</code></a></li> <li><a href="https://github.com/docker/metadata-action/blob/master/#tag-edge--tag-edge-branch"><code>tag-edge</code> / <code>tag-edge-branch</code></a></li> <li><a href="https://github.com/docker/metadata-action/blob/master/#tag-semver"><code>tag-semver</code></a></li> <li><a href="https://github.com/docker/metadata-action/blob/master/#tag-match--tag-match-group"><code>tag-match</code> / <code>tag-match-group</code></a></li> <li><a href="https://github.com/docker/metadata-action/blob/master/#tag-latest"><code>tag-latest</code></a></li> <li><a href="https://github.com/docker/metadata-action/blob/master/#tag-schedule"><code>tag-schedule</code></a></li> <li><a href="https://github.com/docker/metadata-action/blob/master/#tag-custom--tag-custom-only"><code>tag-custom</code> / <code>tag-custom-only</code></a></li> <li><a href="https://github.com/docker/metadata-action/blob/master/#label-custom"><code>label-custom</code></a></li> </ul> </li> <li><a href="https://github.com/docker/metadata-action/blob/master/#basic-workflow">Basic workflow</a></li> <li><a href="https://github.com/docker/metadata-action/blob/master/#semver-workflow">Semver workflow</a></li> </ul> <h3>inputs</h3> <table> <thead> <tr> <th>New</th> <th>Unchanged</th> <th>Removed</th> </tr> </thead> <tbody> <tr> <td><code>tags</code></td> <td><code>images</code></td> <td><code>tag-sha</code></td> </tr> <tr> <td><code>flavor</code></td> <td><code>sep-tags</code></td> <td><code>tag-edge</code></td> </tr> <tr> <td><code>labels</code></td> <td><code>sep-labels</code></td> <td><code>tag-edge-branch</code></td> </tr> <tr> <td></td> <td></td> <td><code>tag-semver</code></td> </tr> <tr> <td></td> <td></td> <td><code>tag-match</code></td> </tr> <tr> <td></td> <td></td> <td><code>tag-match-group</code></td> </tr> <tr> <td></td> <td></td> <td><code>tag-latest</code></td> </tr> <tr> <td></td> <td></td> <td><code>tag-schedule</code></td> </tr> <tr> <td></td> <td></td> <td><code>tag-custom</code></td> </tr> <tr> <td></td> <td></td> <td><code>tag-custom-only</code></td> </tr> <tr> <td></td> <td></td> <td><code>label-custom</code></td> </tr> </tbody> </table> <h4><code>tag-sha</code></h4> <pre lang="yaml"><code>tags: \| type=sha </code></pre> <h4><code>tag-edge</code> / <code>tag-edge-branch</code></h4> <pre lang="yaml"><code>tags: \| # default branch type=edge </tr></table> </code></pre> </blockquote> <p>... (truncated)</p> </details> <details> <summary>Commits</summary> <ul> <li><a href="`69f6fc9d46`"><code>69f6fc9</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/203">#203</a> from crazy-max/san-fix</li> <li><a href="`2f5b5ae8bf`"><code>2f5b5ae</code></a> Sanitize tag earlier</li> <li><a href="`f206c36955`"><code>f206c36</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/202">#202</a> from crazy-max/v4-prep</li> <li><a href="`a20adfa74e`"><code>a20adfa</code></a> readme: set metadata-action to v4</li> <li><a href="`26b9439ce3`"><code>26b9439</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/201">#201</a> from crazy-max/fix-sanitization</li> <li><a href="`467883f452`"><code>467883f</code></a> Merge pull request <a href="https://github-redirect.dependabot.com/docker/metadata-action/issues/176">#176</a> from crazy-max/node-16</li> <li><a href="`5edf56f2c4`"><code>5edf56f</code></a> Node 16 as default runtime</li> <li><a href="`678218f2be`"><code>678218f</code></a> Note about image name and tag sanitization</li> <li><a href="`e44c1fbe6e`"><code>e44c1fb</code></a> Do not sanitize before pattern matching</li> <li>See full diff in <a href="https://github.com/docker/metadata-action/compare/v3...v4">compare view</a></li> </ul> </details> <br /> [![Dependabot compatibility score](https://dependabot-badges.githubapp.com/badges/compatibility_score?dependency-name=docker/metadata-action&package-manager=github_actions&previous-version=3&new-version=4)](https://docs.github.com/en/github/managing-security-vulnerabilities/about-dependabot-security-updates#about-compatibility-scores) Dependabot will resolve any conflicts with this PR as long as you don't alter it yourself. You can also trigger a rebase manually by commenting ``@dependabot` rebase`. [//]: # (dependabot-automerge-start) [//]: # (dependabot-automerge-end) --- <details> <summary>Dependabot commands and options</summary> <br /> You can trigger Dependabot actions by commenting on this PR: - ``@dependabot` rebase` will rebase this PR - ``@dependabot` recreate` will recreate this PR, overwriting any edits that have been made to it - ``@dependabot` merge` will merge this PR after your CI passes on it - ``@dependabot` squash and merge` will squash and merge this PR after your CI passes on it - ``@dependabot` cancel merge` will cancel a previously requested merge and block automerging - ``@dependabot` reopen` will reopen this PR if it is closed - ``@dependabot` close` will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually - ``@dependabot` ignore this major version` will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this minor version` will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself) - ``@dependabot` ignore this dependency` will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself) </details> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2022-06-29 08:07:10 +00:00
Ryan Russell	28609c4176	chore(auth test): Correct `compute_authorized_search` macro Signed-off-by: Ryan Russell <git@ryanrussell.org>	2022-06-28 18:09:26 -05:00
Ryan Russell	a626cf4c99	chore: test comment readability improvements Signed-off-by: Ryan Russell <git@ryanrussell.org>	2022-06-28 18:09:25 -05:00
Ryan Russell	9b660e1058	chore(routes): correct typo Signed-off-by: Ryan Russell <git@ryanrussell.org>	2022-06-28 18:09:22 -05:00
dependabot[bot]	38b85ec547	Bump docker/setup-buildx-action from 1 to 2 Bumps [docker/setup-buildx-action](https://github.com/docker/setup-buildx-action) from 1 to 2. - [Release notes](https://github.com/docker/setup-buildx-action/releases) - [Commits](https://github.com/docker/setup-buildx-action/compare/v1...v2) --- updated-dependencies: - dependency-name: docker/setup-buildx-action dependency-type: direct:production update-type: version-update:semver-major ... Signed-off-by: dependabot[bot] <support@github.com>	2022-06-28 20:21:43 +00:00
dependabot[bot]	ed185fb636	Bump actions/checkout from 2 to 3 Bumps [actions/checkout](https://github.com/actions/checkout) from 2 to 3. - [Release notes](https://github.com/actions/checkout/releases) - [Changelog](https://github.com/actions/checkout/blob/main/CHANGELOG.md) - [Commits](https://github.com/actions/checkout/compare/v2...v3) --- updated-dependencies: - dependency-name: actions/checkout dependency-type: direct:production update-type: version-update:semver-major ... Signed-off-by: dependabot[bot] <support@github.com>	2022-06-28 20:21:41 +00:00
dependabot[bot]	f7b47b43f4	Bump docker/setup-qemu-action from 1 to 2 Bumps [docker/setup-qemu-action](https://github.com/docker/setup-qemu-action) from 1 to 2. - [Release notes](https://github.com/docker/setup-qemu-action/releases) - [Commits](https://github.com/docker/setup-qemu-action/compare/v1...v2) --- updated-dependencies: - dependency-name: docker/setup-qemu-action dependency-type: direct:production update-type: version-update:semver-major ... Signed-off-by: dependabot[bot] <support@github.com>	2022-06-28 20:21:38 +00:00
dependabot[bot]	8e703fbabe	Bump codecov/codecov-action from 1 to 3 Bumps [codecov/codecov-action](https://github.com/codecov/codecov-action) from 1 to 3. - [Release notes](https://github.com/codecov/codecov-action/releases) - [Changelog](https://github.com/codecov/codecov-action/blob/master/CHANGELOG.md) - [Commits](https://github.com/codecov/codecov-action/compare/v1...v3) --- updated-dependencies: - dependency-name: codecov/codecov-action dependency-type: direct:production update-type: version-update:semver-major ... Signed-off-by: dependabot[bot] <support@github.com>	2022-06-28 20:21:35 +00:00
dependabot[bot]	7ced5c2cc7	Bump docker/metadata-action from 3 to 4 Bumps [docker/metadata-action](https://github.com/docker/metadata-action) from 3 to 4. - [Release notes](https://github.com/docker/metadata-action/releases) - [Upgrade guide](https://github.com/docker/metadata-action/blob/master/UPGRADE.md) - [Commits](https://github.com/docker/metadata-action/compare/v3...v4) --- updated-dependencies: - dependency-name: docker/metadata-action dependency-type: direct:production update-type: version-update:semver-major ... Signed-off-by: dependabot[bot] <support@github.com>	2022-06-28 20:21:32 +00:00
bors[bot]	ef95d1d545	Merge #2561 2561: Add dependabot for GHA r=Kerollmops a=curquiza No panic, I only add dependabot to update our CI, not the rust dependencies. Indeed, no easy command exists for this. If it's too much regular notification, as usual, I will remove it. Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-06-28 20:02:30 +00:00
Clémentine Urquizar	f98c8d7f8b	Add dependabot for GHA	2022-06-28 18:58:23 +02:00
bors[bot]	4862993482	Merge #2525 2525: Auth: Provide all document related permissions for action document.* r=Kerollmops a=janithpet Added a `Action::DocumentsAll` identifier as [suggested](https://github.com/meilisearch/meilisearch/issues/2080#issuecomment-1022952486), along with the other necessary changes in `action.rs`. Inside `store.rs`, added an extra condition in `HeedAuthStore::put_api_key` to append all document related permissions if `key.actions.contains(&DocumentsAll)`. Updated the tests as [suggested](https://github.com/meilisearch/meilisearch/issues/2080#issuecomment-1022952486). I am quite new to Rust, so please let me know if I had made any mistakes; have I written the code in the most idiomatic/efficient way? I am aware that the way I append the document permissions could create duplicates in the `actions` vector, but I am not sure how fix that in a simple way (other than using other dependencies like [itertools](https://github.com/rust-itertools/itertools), for example). ## What does this PR do? Fixes #2080 ## PR checklist Please check if your PR fulfills the following requirements: - [ ] Does this PR fix an existing issue? - [ x] Have you read the contributing guidelines? - [ x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: janithPet <jpetangoda@gmail.com>	2022-06-28 14:02:06 +00:00
bors[bot]	8a64ee0c14	Merge #2557 2557: add more tests on the formatted route r=Kerollmops a=irevoire We had a bunch of tests trying to send array of array in a get request, this is actually not supported thus I updated the tests to only send a single array or a direct string. Also, the real tests that ensure the array of array are well handled are in milli so I don’t think we should lose time trying to « improve » our test surface on this point. Co-authored-by: Irevoire <tamo@meilisearch.com>	2022-06-28 13:41:53 +00:00
Irevoire	05ee2eff01	add more tests on the formatted route	2022-06-28 13:17:55 +02:00
bors[bot]	c6f4775fde	Merge #568 568: Fix not equal filter when field contains both number and strings r=Kerollmops a=GraDKh Related to https://github.com/meilisearch/meilisearch/issues/2516 Looks like the issue should be moved to this repo, but I'm not sure what the right procedure for it. Co-authored-by: Dmytro Gordon <dmytro@bigstream.co>	2022-06-28 08:46:23 +00:00
Dmytro Gordon	3ff03a3f5f	Fix not equal filter when field contains both number and strings	2022-06-27 15:55:17 +03:00
bors[bot]	7ee8855499	Merge #2553 2553: Fix ci binary push r=irevoire a=curquiza Prevent the check of version when releasing a RC for the binary publish. Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-06-27 11:18:40 +00:00
Clémentine Urquizar	7f4fab876d	Add fix to publish-binaries.yml	2022-06-27 13:11:58 +02:00
bors[bot]	688d6f704b	Merge #2549 2549: Fix CI checking version compatibility r=irevoire a=curquiza - Fix CI checks in `if` regarding the `stable` variable created - Fix command to remove prefix `refs/tags/v` from `GITHUB_REF` in `check-release.sh` script - Change from `sh` to `bash` for `check-release.sh` script - Fix error message Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-06-27 08:28:31 +00:00
Clémentine Urquizar	f83188fd60	Fix CI with check-release.sh script	2022-06-24 13:11:04 +02:00
bors[bot]	9e261b996f	Merge #2543 2543: fix all the array on the search get route and improve the tests r=curquiza a=irevoire fix #2527 Co-authored-by: Tamo <tamo@meilisearch.com>	2022-06-23 14:51:36 +00:00
bors[bot]	90b0a4e99f	Merge #2546 2546: Bump milli to 0.31.1 r=irevoire a=Kerollmops Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-06-23 11:07:20 +00:00
Kerollmops	dad86fc3d6	Make the changes necessary to use milli 0.31.1	2022-06-23 10:47:49 +02:00
Kerollmops	7feb15df28	Bump milli to 0.31.1	2022-06-23 10:47:48 +02:00
bors[bot]	bf865f51bb	Merge #2544 2544: Fix content of dump/assets for testing r=curquiza a=loiclec This is just a change to the content of two .dump files used in the integration tests. Those files contained settings with the criterion `desc(fame)`, which is invalid for a v3 or higher dump. The change took place in the updates.json file inside the decompressed .dump files. Instances of `desc(field)` or `asc(field)` were changed to `field:desc` and `field:asc`. The tests were (wrongly) passing because the ranking rules were never parsed. Co-authored-by: Loïc Lecrenier <loic@meilisearch.com>	2022-06-22 15:12:19 +00:00
bors[bot]	83ad1aaf05	Merge #567 567: Bump the milli version to 0.31.1 r=curquiza a=Kerollmops Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-06-22 15:07:03 +00:00
Kerollmops	cc48992e79	Bump the milli version to 0.31.1	2022-06-22 17:05:51 +02:00
bors[bot]	68bb170732	Merge #566 566: Introduce the copy_to_path method on the Index r=irevoire a=Kerollmops Meilisearch needs this method to do snapshots. Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-06-22 14:52:19 +00:00
Kerollmops	238692a8e7	Introduce the copy_to_path method on the Index	2022-06-22 16:49:47 +02:00
Tamo	13f258513f	meilisearch-lib was missing a feature in its cargo.toml	2022-06-22 16:44:16 +02:00
bors[bot]	290a40b7a5	Merge #564 564: Rename the limitedTo parameter into maxTotalHits r=curquiza a=Kerollmops This PR is related to https://github.com/meilisearch/meilisearch/issues/2542, it renames the `limitedTo` parameter into `maxTotalHits`. Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-06-22 13:48:33 +00:00
Loïc Lecrenier	f8aa21bc16	Fix content of dump/assets for testing Some contained settings with the criterion desc(fame), which is invalid for a v3 or higher dump. The change took place in the updates.json file inside the decompressed .dump files. Instances of desc(field) or asc(field) were changed to field:desc and field:asc	2022-06-22 14:51:52 +02:00
bors[bot]	1ffe90bf15	Merge #2530 2530: Check the version in Cargo.toml before publishing r=irevoire a=curquiza Fixes #2079 Also - improves the current docker CI for v0.28.0: the current implementation will make run 2 CI instead of just one for the official release - move the `is-latest-releaes.sh` script, and update the documentation comment - fix version of permissive-json-pointer How to test the script? ``` export GITHUB_REF='refs/tags/v0.28.0' sh .github/scripts/check-release.sh ``` Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com> Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com>	2022-06-22 12:42:29 +00:00
bors[bot]	d546f6f40e	Merge #563 563: Improve the `estimatedNbHits` when a `distinctAttribute` is specified r=irevoire a=Kerollmops This PR is related to https://github.com/meilisearch/meilisearch/issues/2532 but it doesn't fix it entirely. It improves it by computing the excluded documents (the ones with an already-seen distinct value) before stopping the loop, I think it was a mistake and should always have been this way. The reason it doesn't fix the issue is that Meilisearch is lazy, just to be sure not to compute too many things and answer by taking too much time. When we deduplicate the documents by their distinct value we must do it along the water, everytime we see a new document we check that its distinct value of it doesn't collide with an already returned document. The reason we can see the correct result when enough documents are fetched is that we were lucky to see all of the different distinct values possible in the dataset and all of the deduplication was done, no document can be returned. If we wanted to implement that to have a correct `extimatedNbHits` every time we should have done a pass on the whole set of possible distinct values for the distinct attribute and do a big intersection, this could cost a lot of CPU cycles. Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-06-22 12:39:44 +00:00
Tamo	c47369b502	fix all the array on the search get route and improve the tests	2022-06-22 14:33:44 +02:00
Clémentine Urquizar	32c8846514	Rollback 0.0.0 versionning	2022-06-22 12:20:12 +02:00
bors[bot]	38a8d3cae1	Merge #565 565: Bump the milli version to 0.31.0 r=curquiza a=Kerollmops Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-06-22 10:09:41 +00:00
Kerollmops	f5c3b951bc	Bump the milli version to 0.31.0	2022-06-22 12:08:16 +02:00
Kerollmops	d7c248042b	Rename the limitedTo parameter into maxTotalHits	2022-06-22 12:00:48 +02:00
Kerollmops	d2f84a9d9e	Improve the estimatedNbHits when distinct is enabled	2022-06-22 11:39:21 +02:00
Clémentine Urquizar	7490383d4f	Update the not-released version in Cargo.toml files	2022-06-21 19:17:33 +02:00
Clémentine Urquizar	c6ed756dbc	Update script after review	2022-06-21 10:46:32 +02:00
Clémentine Urquizar - curqui	de16de20f4	Update .github/scripts/check-release.sh Co-authored-by: Tamo <tamo@meilisearch.com>	2022-06-21 10:14:24 +02:00
Clémentine Urquizar - curqui	c484d28646	Update .github/scripts/check-release.sh Co-authored-by: Tamo <tamo@meilisearch.com>	2022-06-21 10:14:17 +02:00
Clémentine Urquizar	5318e53248	Move is-latest-release.sh script into the scripts folder	2022-06-21 10:12:00 +02:00
Clémentine Urquizar	06e05cc4f8	Update Docker credentials	2022-06-20 14:39:50 +02:00
bors[bot]	4f547eff02	Merge #560 560: Update version for next release (v0.30.0) r=curquiza a=curquiza Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-06-20 12:37:01 +00:00
bors[bot]	64b833410c	Merge #559 559: Avoid having an ending separator before crop marker r=irevoire a=ManyTheFish related to https://github.com/meilisearch/meilisearch/issues/2528 Co-authored-by: ManyTheFish <many@meilisearch.com>	2022-06-20 11:06:52 +00:00
Clémentine Urquizar	31f749b5d8	Update version for next release (v0.30.0)	2022-06-20 12:09:57 +02:00
bors[bot]	ba839a909f	Merge #2537 2537: Change information place regarding release assets r=Kerollmops a=curquiza Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-06-20 08:40:25 +00:00
Clémentine Urquizar	8b98303191	Change information place regarding release assets	2022-06-20 10:36:34 +02:00
Clémentine Urquizar	2dde6fadb4	Check the version in Cargo.toml before publishing	2022-06-20 10:22:01 +02:00
ManyTheFish	a0ab90a4d7	Avoid having an ending separator before crop marker	2022-06-16 18:23:57 +02:00
bors[bot]	eb8d53a915	Merge #2529 2529: Improve docker CI: push vX.Y tag (without patch) to DockerHub (for v0.28.0) r=curquiza a=curquiza Bringing a commit from main ([`5ae5b06`](`5ae5b06018`)) into release-v0.28.0 to have this change already applied for v0.28.0 Following the one merged on main: https://github.com/meilisearch/meilisearch/pull/2521 Co-authored-by: Janith Petangoda <22471198+janithpet@users.noreply.github.com>	2022-06-16 15:38:25 +00:00
Janith Petangoda	10f3150150	Improve docker CI: push `vX.Y` tag (without patch) to DockerHub (#2507 ) * Create a docker tag without patch version if git tag has 0 patch version. * Create Docker tag without patch number if git tag follows v<number>.<number>.<number> Add minor changes on CI	2022-06-16 17:14:09 +02:00
bors[bot]	54cd9976d7	Merge #2521 2521: Improve docker CI: push `vX.Y` tag (without patch) to DockerHub (#2507) r=curquiza a=curquiza Following https://github.com/meilisearch/meilisearch/pull/2507 Fixes #2497 Co-authored-by: Janith Petangoda <22471198+janithpet@users.noreply.github.com>	2022-06-16 13:28:27 +00:00
bors[bot]	a59ae19842	Merge #558 558: Deletion benchmarks r=ManyTheFish a=ManyTheFish Add benchmarks on the deletion and start rethinking benchmark names. Co-authored-by: ManyTheFish <many@meilisearch.com>	2022-06-16 09:34:37 +00:00
ManyTheFish	2652310f2a	Change delete benchmark names	2022-06-16 10:32:58 +02:00
ManyTheFish	adbb0ff318	Add deletion benchmarks	2022-06-16 10:17:58 +02:00
Janith Petangoda	5ae5b06018	Improve docker CI: push `vX.Y` tag (without patch) to DockerHub (#2507 ) * Create a docker tag without patch version if git tag has 0 patch version. * Create Docker tag without patch number if git tag follows v<number>.<number>.<number> Add minor changes on CI	2022-06-16 09:51:20 +02:00
janithPet	6f910f89eb	Ran formatter on the code.	2022-06-15 22:23:38 +01:00
bors[bot]	b455c3897f	Merge #2524 2524: Add the specific routes for the pagination and faceting settings r=curquiza a=Kerollmops Fixes #2522 Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-06-15 17:01:52 +00:00
janithPet	9a8fb6c55a	Updated actions identifiers to be in a more pleasing order	2022-06-15 17:27:41 +01:00
janithPet	4016161035	Provide all document related permissions for action document.*	2022-06-15 16:11:39 +01:00
Kerollmops	9d692ba1c6	Add more tests for the pagination and faceting subsettings routes	2022-06-15 15:53:43 +02:00
Kerollmops	22e1ac969a	Add specific routes for the pagination and faceting settings	2022-06-15 15:27:06 +02:00
bors[bot]	3340af1ba9	Merge #2517 2517: Fix typos in `tasks/.rs` r=MarinPostma a=ryanrussell Signed-off-by: Ryan Russell <git@ryanrussell.org> ## What does this PR do? Readability in `tasks/.rs` files. ## PR checklist - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Co-authored-by: Ryan Russell <git@ryanrussell.org>	2022-06-15 09:27:21 +00:00
bors[bot]	49d509936b	Merge #2520 2520: Use nightly for cargo fmt in CI (for `release-v0.28.0` branch) r=Kerollmops a=curquiza Following https://github.com/meilisearch/meilisearch/pull/2519 Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-06-15 09:08:40 +00:00
bors[bot]	e46b853fbf	Merge #2519 2519: Use nightly for cargo fmt in CI r=Kerollmops a=curquiza Rustfmt CI currently does not pass on main with nightly, see https://github.com/meilisearch/meilisearch/pull/2517#issuecomment-1156138216 Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-06-15 08:35:46 +00:00
Clémentine Urquizar	44e004d895	Use nightly for cargo fmt in CI	2022-06-15 10:33:03 +02:00
Ryan Russell	71bf9b5b9b	docs: Readability improvements in `tasks/` Signed-off-by: Ryan Russell <git@ryanrussell.org>	2022-06-14 20:38:45 -05:00
bors[bot]	0a5d1a445e	Merge #554 554: Enhance tests for soft deletetion r=irevoire a=ManyTheFish #### tests: (skip in changelog) - [x] placeholder search shouldn’t return soft deleted - [x] search shouldn’t return soft deleted - [x] filtered placeholder search shouldn’t return soft deleted - [x] geo-filtered placeholder search shouldn’t return soft deleted - [x] documents list/get shouldn’t return soft deleted - [x] stats shouldn’t count soft deleted #### other: (API breaking) - [x] ensure that Index methods are not bypassed by Meilisearch Poke `@irevoire,` we may merge this into your branch. Co-authored-by: ManyTheFish <many@meilisearch.com>	2022-06-14 09:49:37 +00:00
ManyTheFish	447195a27a	Replace format by to_string	2022-06-14 10:32:44 +02:00
ManyTheFish	177154828c	Extends deletion tests	2022-06-13 17:34:16 +02:00
ManyTheFish	0d1d354052	Ensure that Index methods are not bypassed by Meilisearch	2022-06-13 17:34:11 +02:00
bors[bot]	053071d866	Merge #2502 2502: test dump v5 r=ManyTheFish a=MarinPostma this PR adds a test for dump v5 Co-authored-by: ad hoc <postma.marin@protonmail.com>	2022-06-13 10:11:22 +00:00
bors[bot]	0f6e650ba2	Merge #2508 2508: docs: Readability improvements in `download-latest.sh` and `src/analytics/.rs` r=Kerollmops a=ryanrussell # Pull Request ## What does this PR do? Improves readability in `download-latest.sh` and in the `src/analytics/.rs` files ## PR checklist - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Ryan Russell <git@ryanrussell.org>	2022-06-13 08:18:17 +00:00
ad hoc	97daea5a66	ignore dump test on windows	2022-06-13 09:16:36 +01:00
Ryan Russell	8990b12609	docs: Readability fixes in `src/analytics/.rs` files Signed-off-by: Ryan Russell <git@ryanrussell.org>	2022-06-11 10:21:05 -05:00
Ryan Russell	07a35c6445	docs: Bash comment readability fixes Signed-off-by: Ryan Russell <git@ryanrussell.org>	2022-06-11 10:20:28 -05:00
bors[bot]	4fc73195e6	Merge #2466 2466: index resolver tests r=MarinPostma a=MarinPostma add more index resolver tests depends on #2455 followup #2453 Co-authored-by: ad hoc <postma.marin@protonmail.com>	2022-06-09 17:50:01 +00:00
ad hoc	1425d62a31	test dump v5	2022-06-09 18:35:17 +02:00
bors[bot]	87d4bf672c	Merge #2500 2500: Bump milli r=Kerollmops a=irevoire Fix #2380 Co-authored-by: Tamo <tamo@meilisearch.com>	2022-06-09 16:35:00 +00:00
Tamo	2063fbd985	chore: bump milli	2022-06-09 18:34:03 +02:00
bors[bot]	de356061db	Merge #2414 2414: Improve index uid validation upon API key creation r=Kerollmops a=pierre-l - ~Use an IndexUid newtype to enforce stronger constraints~ - ~`cargo update -p vergen`~ (`rustup update` was the proper fix for this) - Add a new `meilisearch_types` crate - Move `meilisearch_error` to `meilisearch_types::error` - Move `meilisearch_lib::index_resolver::IndexUid` to `meilisearch_types::index_uid` - Add a new `InvalidIndexUid` error in `meilisearch_types::index_uid` - Move `meilisearch_http::routes::StarOr` to `meilisearch_types::star_or` - Use the `IndexUid` and `StarOr` in `meilisearch_auth::Key` Fixes #2158 Co-authored-by: pierre-l <pierre.larger@gmail.com>	2022-06-09 15:41:51 +00:00
ad hoc	9a6841c7ce	remove unused test	2022-06-09 17:10:56 +02:00
ad hoc	354f7fb2bf	test index_update	2022-06-09 17:10:55 +02:00
ad hoc	0333bad057	delete documents test	2022-06-09 17:10:55 +02:00
ad hoc	0e416b4bcd	delete index test	2022-06-09 17:10:55 +02:00
bors[bot]	f1d848bb9a	Merge #552 552: Fix escaped quotes in filter r=Kerollmops a=irevoire Will fix https://github.com/meilisearch/meilisearch/issues/2380 The issue was that in the evaluation of the filter, I was using the deref implementation instead of calling the `value` method of my token. To avoid the problem happening again, I removed the deref implementation; now, you need to either call the `lexeme` or the `value` methods but can't rely on a « default » implementation to get a string out of a token. Co-authored-by: Tamo <tamo@meilisearch.com>	2022-06-09 14:56:44 +00:00
Tamo	676187ba43	bump milli version	2022-06-09 16:53:32 +02:00
bors[bot]	20dd259f23	Merge #2455 2455: refactor perform task r=curquiza a=MarinPostma Refactor some index resolver functions to make duties more consistent, and code easier to test Co-authored-by: ad hoc <postma.marin@protonmail.com>	2022-06-09 14:49:59 +00:00
bors[bot]	18095fa4e1	Merge #2498 2498: Update CONTRIBUTING.md to add the release process r=Kerollmops a=curquiza Add release process to the contributing Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-06-09 14:23:49 +00:00
pierre-l	b8745420da	Use the `IndexUid` and `StarOr` in `meilisearch_auth::Key` Move `meilisearch_http::routes::StarOr` to `meilisearch_types::star_or` Fixes #2158	2022-06-09 16:14:15 +02:00
pierre-l	36cb09eb25	Add a new `meilisearch_types` crate Move `meilisearch_error` to `meilisearch_types::error` Move `meilisearch_lib::index_resolver::IndexUid` to `meilisearch_types::index_uid` Add a new `InvalidIndexUid` error in `meilisearch_types::index_uid`	2022-06-09 16:14:13 +02:00
Tamo	90afde435b	fix escaped quotes in filter	2022-06-09 16:03:49 +02:00
ad hoc	8fc3b7d3b0	refactor process_document_addition_batch	2022-06-09 14:59:20 +02:00
ad hoc	64e3096790	process_task updates task events	2022-06-09 14:59:20 +02:00
ad hoc	b594d49def	add IndexResolver BatchHandler tests	2022-06-09 14:59:19 +02:00
ad hoc	fbba67fbe9	add mocker to IndexResolver	2022-06-09 14:59:19 +02:00
Clémentine Urquizar	232b2baaa3	Update TOC	2022-06-09 14:49:49 +02:00
Clémentine Urquizar	02c5c193a2	Update CONTRIBUTING.md to add the release process	2022-06-09 14:48:50 +02:00
bors[bot]	b9b32d65a8	Merge #2494 2494: Introduce the new faceting and pagination settings r=ManyTheFish a=Kerollmops This PR introduces two new settings following the newly created spec https://github.com/meilisearch/specifications/pull/157: - The `faceting.max_values_per_facet` one describes the maximum number of values (each with a count) associated with a value in a facet distribution query. - The `pagination.limited_to` one describes the maximum number of documents that a search query can ever return. Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-06-09 12:09:21 +00:00
bors[bot]	680606fd82	Merge #2491 2491: Improve Docker CIs r=curquiza a=curquiza Close #1901 Continuing work of - [x] Merge two docker CI in one (#2477). The latest tag is still only push for the official release. - [x] Add cron job in GHA to run the CI (without publishing the image) every day. This will check the docker build is working. Co-authored-by: Lawrence Chou <lawrencechou1024@gmail.com> Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com> Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-06-09 11:26:12 +00:00
Clémentine Urquizar	4a494ad2fa	Add schedule to the CI	2022-06-09 11:50:20 +02:00
bors[bot]	2b2e571c76	Merge #2460 2460: Create custom error types for `TaskType`, `TaskStatus`, and `IndexUid` r=Kerollmops a=walterbm # Pull Request ## What does this PR do? Fixes #2443 by making the following changes: - Add custom `TaskTypeError` for `TaskType::from_str` - Add custom `TaskStatusError` for `TaskStatus::from_str` - Add custom `IndexUidFormatError` for `IndexUid::from_str` - Implement `From<IndexUidFormatError> for IndexResolverError` to convert between errors - Replace all usages of `IndexUid::new` with `IndexUid::from_str` - NOTE I am relatively new to Rust and I struggled a lot with this final part. This PR ended up with a messy error conversion which does not seem ideal. Please let me know if you have any suggestions for how to make this better and I'll be happy to make any updates! ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: walter <walter.beller.morales@gmail.com>	2022-06-09 09:10:28 +00:00
Kerollmops	6f0d3472b1	Change the test for the new pagination.limited_to setting	2022-06-09 10:56:43 +02:00
Kerollmops	5cd13cc303	Add a test to validate the faceting.max_values_per_facet setting	2022-06-09 10:56:42 +02:00
Kerollmops	1e3dcbea3f	Plug the pagination.limited_to setting	2022-06-09 10:56:42 +02:00
Kerollmops	b96399d24b	Plug the faceting.max_values_per_facet setting	2022-06-09 10:56:42 +02:00
Kerollmops	5450b5ced3	Add the faceting.max_values_per_facet setting	2022-06-09 10:54:32 +02:00
Kerollmops	c924614527	Bump milli to 0.29.2	2022-06-09 10:54:28 +02:00
bors[bot]	19d44142a1	Merge #550 550: Add the two new pagination and faceting settings r=ManyTheFish a=Kerollmops This PR adds two new settings in the database, those settings are described [in this spec](https://github.com/meilisearch/specifications/pull/157). Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-06-09 08:16:01 +00:00
walter	96d4fd54bb	Change the index uid format check for better legibility	2022-06-08 19:58:47 -04:00
walter	3e5d6be86b	Rename TaskType::from_str parameter to 'type_'	2022-06-08 19:57:45 -04:00
walter	2b944ecd89	Remove IndexUid::new and replace with IndexUid::from_str	2022-06-08 19:56:01 -04:00
bors[bot]	db42268888	Merge #2473 2473: fix blocking in dumps r=irevoire a=MarinPostma This PR fixes two blocking calls in the dump process. Co-authored-by: ad hoc <postma.marin@protonmail.com>	2022-06-08 17:14:46 +00:00
ad hoc	108b3520de	fix blocking auth controller dump	2022-06-08 18:19:29 +02:00
Kerollmops	445d5474cc	Add the pagination_limited_to setting to the database	2022-06-08 18:14:27 +02:00
bors[bot]	f64b824c45	Merge #2493 2493: Update version for next release (v0.28.0) r=Kerollmops a=curquiza Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-06-08 16:02:03 +00:00
Clémentine Urquizar	fc4990b968	Update version for next release (v0.28.0)	2022-06-08 17:59:18 +02:00
Kerollmops	69931e50d2	Add the max_values_by_facet setting to the database	2022-06-08 17:54:56 +02:00
Clémentine Urquizar - curqui	39a1dcb32c	Update .github/workflows/publish-docker.yml	2022-06-08 17:27:03 +02:00
Lawrence Chou	afcc493480	Merge publish-docker-latest.yml & publish-docker-tag.yml (#2477 ) close #1901	2022-06-08 17:17:20 +02:00
Kerollmops	52a494bd3b	Add the new pagination.limited_to and faceting.max_values_per_facet settings	2022-06-08 17:15:36 +02:00
bors[bot]	9580b9de79	Merge #549 549: Bump the version to 0.29.2 r=curquiza a=Kerollmops Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-06-08 14:29:47 +00:00
bors[bot]	6171f17f1d	Merge #2468 2468: Update milli 0.29 r=Kerollmops a=ManyTheFish - [x] Update milli to 0.29 - [x] Integrate charabia - [x] Set disabled_words to default when Index::exact_words returns None - [x] Fix ranking rules integration test fixes #2375 fixes #2144 fixes #2417 fixes #2407 Co-authored-by: ManyTheFish <many@meilisearch.com>	2022-06-08 14:29:20 +00:00
bors[bot]	a762d7f462	Merge #548 548: Setup the new limits on the number of facet values to return r=ManyTheFish a=Kerollmops This PR implements the early draft of the new spec (waiting for it) specifying how the new facet limit feature should work and which limit we apply to the number of facet values to return by facet. Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-06-08 14:00:33 +00:00
Kerollmops	56ee9cc21f	Bump the version to 0.29.2	2022-06-08 16:00:06 +02:00
Kerollmops	2a505503b3	Change the number of facet values returned by default to 100	2022-06-08 15:58:57 +02:00
Kerollmops	bae4007447	Remove the hard limit on the number of facet values returned	2022-06-08 15:58:57 +02:00
ManyTheFish	55169ff914	Fix test get_document_s_nested_attributes_to_retrieve	2022-06-08 15:56:06 +02:00
ManyTheFish	0a16f71563	Increase wait_task wainting time	2022-06-08 15:56:03 +02:00
bors[bot]	bd280f75e7	Merge #2487 #2489 2487: Feat(auth): Use hmac instead of sha256 to derivate API keys from master key r=MarinPostma a=ManyTheFish Wrap sha256 in HMAC to derivate new API keys from the master key. 2489: Fix(auth): Authorization test were not testing keys unrestricted on index r=Kerollmops a=ManyTheFish Co-authored-by: ManyTheFish <many@meilisearch.com>	2022-06-08 13:48:07 +00:00
bors[bot]	617802bae7	Merge #2480 2480: Bring release v0.27.2 changes into stable r=Kerollmops a=curquiza Co-authored-by: Tamo <tamo@meilisearch.com>	2022-06-08 13:08:26 +00:00
ManyTheFish	1a7631c807	Hash master_key before passing it to HMAC	2022-06-08 14:54:47 +02:00
ManyTheFish	17f30c2b2d	Fix(auth): Authorization test were not testing keys unrestricted on index	2022-06-08 14:52:32 +02:00
ManyTheFish	09938c9b6f	Patch ranking rules error test	2022-06-08 14:38:09 +02:00
ManyTheFish	f5306eb5b0	Set disabled_words to default when Index::exact_words returns None	2022-06-08 14:38:09 +02:00
ManyTheFish	173eea06e1	Replace old tokenizer by charabia	2022-06-08 14:38:09 +02:00
ManyTheFish	8d09772334	Update milli	2022-06-08 14:38:05 +02:00
ManyTheFish	987a7f8926	Wrap sha256 in HMAC instead of directly use sha256	2022-06-08 14:25:12 +02:00
bors[bot]	0928f3d41c	Merge #2453 2453: test index resolver r=MarinPostma a=MarinPostma add some tests to the `IndexResolver` implementation of `BatchHandler` Co-authored-by: ad hoc <postma.marin@protonmail.com>	2022-06-08 11:05:30 +00:00
bors[bot]	7313d6c533	Merge #547 547: Update version for next release (v0.29.1) r=Kerollmops a=curquiza A new milli version will be released once this PR is merged https://github.com/meilisearch/milli/pull/543 Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-06-08 10:20:24 +00:00
bors[bot]	306d2f37ff	Merge #543 543: Fix wrong internal ids assignments r=irevoire a=irevoire Fix https://github.com/meilisearch/meilisearch/issues/2470 Co-authored-by: ad hoc <postma.marin@protonmail.com> Co-authored-by: Tamo <tamo@meilisearch.com>	2022-06-08 09:19:58 +00:00
bors[bot]	09ec8e9fca	Merge #2471 2471: Remove the connection keep-alive timeout r=MarinPostma a=Thearas # Pull Request ## What does this PR do? Fixes <https://github.com/meilisearch/meilisearch-go/issues/221>. Meilisearch has a default connection keep-alive timeout for 5s, which means it will close the connections with idle time >= 5s. This PR set actix-web keep-alive config to `KeepAlive::Os`, let the client and system to decide when to close the connection. ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Co-authored-by: Thearas <thearas850@gmail.com>	2022-06-08 06:59:25 +00:00
Clémentine Urquizar	478dbfa45a	Update version for next release (v0.29.1)	2022-06-07 18:59:33 +02:00
bors[bot]	1968950b0f	Merge #2475 2475: feat(auth): Paginate API keys listing r=irevoire a=ManyTheFish - [x] Update tests - [x] Use Pagination helpers to paginate API keys (thanks to `@irevoire)` fixes #2442 Co-authored-by: ManyTheFish <many@meilisearch.com>	2022-06-07 15:42:02 +00:00
ManyTheFish	6ffa222218	feat(auth): Paginate API keys listing - [x] Update tests - [x] Use Pagination helpers to paginate API keys fixes #2442	2022-06-07 17:37:48 +02:00
bors[bot]	79e67df73d	Merge #2474 2474: feat(auth): remove `dumps.get` action from keys r=ManyTheFish a=ManyTheFish - [x] Update tests - [x] Remove `dumps.get` action related to: #2430 Co-authored-by: ManyTheFish <many@meilisearch.com>	2022-06-07 15:03:41 +00:00
bors[bot]	72be296852	Merge #2476 2476: fix(test): Windows index pagination test r=ManyTheFish a=ManyTheFish The test `index::get_index::get_and_paginate_indexes` fails on every PRs during `Tests on windows-latest` CI job. - reduce the default index size in tests. - wait for each task instead of waiting for the last one at the end. Co-authored-by: ManyTheFish <many@meilisearch.com>	2022-06-07 14:17:01 +00:00
Tamo	d0aaa7ff00	Fix wrong internal ids assignments	2022-06-07 15:49:33 +02:00
ad hoc	31776fdc3f	add failing test	2022-06-07 15:49:33 +02:00
ManyTheFish	a7bff35e49	fix(test): Reduce default index size in tests	2022-06-07 15:16:34 +02:00
ManyTheFish	3b01ed4fe8	feat(auth): remove `dumps.get` action from keys	2022-06-07 10:49:28 +02:00
ad hoc	cbd27d313c	fix blocking writing of meta file in dump	2022-06-07 10:07:40 +02:00
ad hoc	6ac8675c6d	add IndexResolver BatchHandler tests	2022-06-07 09:33:57 +02:00
ad hoc	df61ca9cae	add mocker to IndexResolver	2022-06-07 09:33:57 +02:00
ad hoc	bbd685af5e	move IndexResolver to real module	2022-06-07 09:33:56 +02:00
Thearas	9b9cbc815b	fmt	2022-06-07 03:50:39 +08:00
Thearas	fd11903920	remove the connection timeout	2022-06-07 03:38:23 +08:00
bors[bot]	c3003065e8	Merge #2464 2464: Simplify the `star_or` function usage r=ManyTheFish a=Kerollmops This PR simplifies the usage of the `star_or` function that was originally introduced in #2399. The `serde-cs` https://github.com/naughie/serde-cs/pull/1 PR was merged and was implementing the `IntoIterator` trait on the `CS` types, which makes it possible to directly collect without converting a `CS` into the inner type (vec). Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-06-06 12:47:27 +00:00
bors[bot]	c6ce3452cf	Merge #2459 2459: Fix typo in codebase comments r=Kerollmops a=ryanrussell # Pull Request ## PR checklist - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Co-authored-by: Ryan Russell <git@ryanrussell.org>	2022-06-06 09:01:07 +00:00
Kerollmops	e5b760c59a	Fix the segment analytics tests	2022-06-06 10:44:46 +02:00
Kerollmops	277a0a7967	Bump serde-cs to simplify our usage of the star_or function	2022-06-06 10:17:33 +02:00
Kerollmops	64b5b2e1f8	Use serde-cs::CS with StarOr to reduce the logic duplication	2022-06-06 10:06:00 +02:00
Kerollmops	10d3b367dc	Simplify the const default values	2022-06-06 10:06:00 +02:00
walter	ba55905377	Add custom IndexUidFormatError for IndexUid	2022-06-05 02:26:48 -04:00
walter	0e7e16ae72	Add custom TaskStatusError for TaskStatus	2022-06-05 00:51:08 -04:00
walter	80c156df3f	Add custom TaskTypeError for TaskType	2022-06-05 00:51:08 -04:00
Ryan Russell	4b6c3e72ff	Improve Lib Readability Signed-off-by: Ryan Russell <git@ryanrussell.org>	2022-06-04 21:38:04 -05:00
Ryan Russell	3e46543060	Improve Store Readability Signed-off-by: Ryan Russell <git@ryanrussell.org>	2022-06-04 20:42:53 -05:00
bors[bot]	05ae6dbfa4	Merge #541 541: Update version for next release (v0.29.0) r=ManyTheFish a=curquiza Need to update the version since #540 was merged and breaking Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-06-02 16:53:28 +00:00
bors[bot]	78f76c841d	Merge #542 542: Refactor matching word r=Kerollmops a=ManyTheFish Simplify MatchingWords API Co-authored-by: ManyTheFish <many@meilisearch.com>	2022-06-02 16:23:41 +00:00
ManyTheFish	d212dc6b8b	Remove useless newline	2022-06-02 18:22:56 +02:00
ManyTheFish	a5c790bf4b	Update http-ui	2022-06-02 18:15:36 +02:00
Clémentine Urquizar	6ce1c6487a	Update version for next release (v0.29.0)	2022-06-02 18:07:55 +02:00
ManyTheFish	727d663f28	Update benchmarks	2022-06-02 18:07:10 +02:00
ManyTheFish	7aabe42ae0	Refactor matching words	2022-06-02 17:59:04 +02:00
bors[bot]	dd186533f0	Merge #540 540: Integrate charabia r=Kerollmops a=ManyTheFish related to https://github.com/meilisearch/meilisearch/issues/2375 related to https://github.com/meilisearch/meilisearch/issues/2144 related to https://github.com/meilisearch/meilisearch/issues/2417 Co-authored-by: ManyTheFish <many@meilisearch.com>	2022-06-02 15:34:33 +00:00
ManyTheFish	4dd7b20c32	Update benchmarks	2022-06-02 17:33:25 +02:00
bors[bot]	b83455f345	Merge #2454 2454: Unify the pagination of the index and documents route behind a common type r=curquiza a=irevoire `@MarinPostma` wdyt of keeping the `auto_paginate_sized` until we implement the pagination on every route that needs it just to see if it could be useful to something else Co-authored-by: Tamo <tamo@meilisearch.com>	2022-06-02 15:01:43 +00:00
ManyTheFish	4dd3675d2b	Update http-ui	2022-06-02 16:59:11 +02:00
ManyTheFish	86ac8568e6	Use Charabia in milli	2022-06-02 16:59:11 +02:00
ManyTheFish	192e024ada	Add Charabia in Cargo.toml	2022-06-02 16:59:07 +02:00
bors[bot]	953a209f02	Merge #2447 2447: move index uid in task content r=Kerollmops a=MarinPostma this pr moves the index_uid from the `Task` to the `TaskContent`. This is because the task can now have content that do not target a particular index. Co-authored-by: ad hoc <postma.marin@protonmail.com>	2022-06-02 13:54:09 +00:00
ad hoc	0c5352fc22	move index_uid from task to task_content	2022-06-02 15:30:35 +02:00
bors[bot]	8ac8fcb0c9	Merge #2433 2433: Fix the documents route r=Kerollmops a=irevoire fix #2428 Co-authored-by: Irevoire <tamo@meilisearch.com>	2022-06-02 13:25:34 +00:00
Irevoire	4667c9fe1a	fix(http): Fix the query parameter in the Documents route	2022-06-02 14:10:44 +02:00
Tamo	12b5eabd5d	chore(http): unify the pagination of the index and documents route behind a common type	2022-06-02 14:06:56 +02:00
bors[bot]	cf2d8de48a	Merge #2452 2452: Change http verbs r=ManyTheFish a=Kerollmops This PR fixes #2419 by updating the HTTP verbs used to update the settings and every single setting parameter. - [x] `PATCH /indexes/{indexUid}` instead of `PUT` - [x] `PATCH /indexes/{indexUid}/settings` instead of `POST` - [x] `PATCH /indexes/{indexUid}/settings/typo-tolerance` instead of `POST` - [x] `PUT /indexes/{indexUid}/settings/displayed-attributes` instead of `POST` - [x] `PUT /indexes/{indexUid}/settings/distinct-attribute` instead of `POST` - [x] `PUT /indexes/{indexUid}/settings/filterable-attributes` instead of `POST` - [x] `PUT /indexes/{indexUid}/settings/ranking-rules` instead of `POST` - [x] `PUT /indexes/{indexUid}/settings/searchable-attributes` instead of `POST` - [x] `PUT /indexes/{indexUid}/settings/sortable-attributes` instead of `POST` - [x] `PUT /indexes/{indexUid}/settings/stop-words` instead of `POST` - [x] `PUT /indexes/{indexUid}/settings/synonyms` instead of `POST` Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-06-02 11:46:17 +00:00
Kerollmops	419922e475	Make clippy happy	2022-06-02 13:38:23 +02:00
bors[bot]	c9cd1738a5	Merge #2445 2445: Seek-based tasks list r=Kerollmops a=Kerollmops This PR implements the seek-based pagination for the tasks list following [the spec](https://github.com/meilisearch/specifications/pull/115). Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-06-02 10:25:54 +00:00
Kerollmops	0258659278	Fix the get_settings tests	2022-06-02 12:24:27 +02:00
Kerollmops	ce37f53a16	Add typo-tolerance to the authorization tests	2022-06-02 12:17:53 +02:00
Kerollmops	bcb51905d7	Fix the authorization tests	2022-06-02 12:16:46 +02:00
Kerollmops	10a71fdb10	Update the /indexes/{indexUid}/settings/* verbs by adding a macro parameter	2022-06-02 11:55:47 +02:00
Kerollmops	f8d3f739ad	Update the /indexes/{indexUid}/settings verb from POST to PATCH	2022-06-02 11:55:47 +02:00
Kerollmops	bb405aa729	Update the /indexes/{indexUid} verb from PUT to PATCH	2022-06-02 11:55:47 +02:00
bors[bot]	7e3d5ebc8e	Merge #2451 2451: feat(API-keys): Change immutable_field error message r=Kerollmops a=ManyTheFish Change the immutable_field error message to fit the recent changes in the spec: `aa0a148ee3..84a9baff68` Co-authored-by: ManyTheFish <many@meilisearch.com>	2022-06-02 09:26:36 +00:00
Kerollmops	dfce9ba468	Apply suggestions	2022-06-02 11:26:12 +02:00
ManyTheFish	9eea142e2b	feat(API-keys): Change immutable_field error message Change the immutable_field error message to fit the recent changes in the spec: `aa0a148ee3..84a9baff68`	2022-06-02 11:11:07 +02:00
bors[bot]	8b8c3e32f0	Merge #2450 2450: Bump the dependencies r=ManyTheFish a=Kerollmops In order to use [the latest version of grenad](https://docs.rs/grenad) I bump the dependencies here. We also use the latest versions of all our other dependencies now. Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-06-02 08:53:12 +00:00
bors[bot]	08d72e32a4	Merge #2438 2438: Refine keys api r=ManyTheFish a=ManyTheFish waiting for #2410 and #2444 to be merged. fix #2369 Co-authored-by: ManyTheFish <many@meilisearch.com>	2022-06-02 08:32:35 +00:00
Kerollmops	ac9e7bdbe3	Fix a test that was depending on the speed of the CPU	2022-06-02 10:21:19 +02:00
bors[bot]	ac6df0df57	Merge #539 539: Update version to v0.28.1 r=Kerollmops a=curquiza Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-06-01 16:40:12 +00:00
Clémentine Urquizar	c19c17eddb	Update version to v0.28.1	2022-06-01 18:31:02 +02:00
ManyTheFish	4512eed8f5	Fix PR comments	2022-06-01 18:06:20 +02:00
Kerollmops	e769043576	Bump the dependencies	2022-06-01 17:59:05 +02:00
Kerollmops	df721b2e9e	Scheduler must not reverse the order of the fetched tasks	2022-06-01 17:16:15 +02:00
Kerollmops	0656df3a6d	Fix the dumps tests	2022-06-01 17:14:13 +02:00
bors[bot]	74d1914a64	Merge #535 535: Reintroduce the max values by facet limit r=ManyTheFish a=Kerollmops This PR reintroduces the max values by facet limit this is related to https://github.com/meilisearch/meilisearch/issues/2349. ~I would like some help in deciding on whether I keep the default 100 max values in milli and set up the `FacetDistribution` settings in Meilisearch to use 1000 as the new value, I expose the `max_values_by_facet` for this purpose.~ I changed the default value to 1000 and the max to 10000, thank you `@ManyTheFish` for the help! Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-06-01 14:30:50 +00:00
ManyTheFish	7652295d2c	Encode key in base64 instead of hexa	2022-06-01 16:17:47 +02:00
ManyTheFish	94b32cce01	Patch errors	2022-06-01 16:17:47 +02:00
ManyTheFish	b2e2dc8558	Re-authorize master_key to access to all routes	2022-06-01 16:17:47 +02:00
ManyTheFish	1816db8c1f	Move dump v4 patcher into v4.rs	2022-06-01 16:17:43 +02:00
ManyTheFish	c295924ea2	Patch tests	2022-06-01 16:08:42 +02:00
ManyTheFish	1f62e83267	Remove error_add_api_key_invalid_index_uid_format	2022-06-01 16:08:42 +02:00
ManyTheFish	b3c8915702	Make small changes and renaming	2022-06-01 16:08:42 +02:00
ManyTheFish	151f494110	Use Stream Deserializer to load dumps	2022-06-01 16:08:42 +02:00
ManyTheFish	96152a3d32	Change default API keys names and descriptions	2022-06-01 16:08:42 +02:00
ManyTheFish	84f52ac175	Add v4 feature to uuid	2022-06-01 16:08:42 +02:00
ManyTheFish	70916d6596	Patch dump v4	2022-06-01 16:08:42 +02:00
ManyTheFish	b9a79eb858	Change apiKeyPrefix to apiKeyUid	2022-06-01 16:07:44 +02:00
ManyTheFish	a57b2d9538	Restrict master key access to /keys routes	2022-06-01 16:07:44 +02:00
ManyTheFish	34c8888f56	Add keys actions	2022-06-01 16:07:44 +02:00
ManyTheFish	d54643455c	Make PATCH only modify name, description, and updated_at fields	2022-06-01 16:07:44 +02:00
ManyTheFish	96a5791e39	Add uid and name fields in keys	2022-06-01 16:07:44 +02:00
ManyTheFish	e2c204cf86	Update tests to fit to the new requirements	2022-06-01 16:07:44 +02:00
Kerollmops	d80e8b64af	Align the tasks route API to the new spec	2022-06-01 15:30:39 +02:00
Kerollmops	c11d21879a	Introduce tasks limit and after to the tasks route	2022-06-01 13:26:36 +02:00
bors[bot]	d6dd234914	Merge #2434 2434: Update docker volume path r=curquiza a=0x0x1 Currently, the args `$(pwd)/data.ms:/data.ms` cannot share data from the container, makes docker volume same as the [Dockerfile](`67b6f4340a/Dockerfile (L40)`) to fixing it. Co-authored-by: 0x0x1 <101086451+0x0x1@users.noreply.github.com> Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com>	2022-06-01 10:31:27 +00:00
Clémentine Urquizar - curqui	4970525541	Update README.md Co-authored-by: Tamo <irevoire@protonmail.ch>	2022-06-01 12:29:36 +02:00
Kerollmops	461b91fd13	Introduce the fetch_unfinished_tasks function to fetch tasks	2022-06-01 12:09:52 +02:00
Kerollmops	004c8b6be3	Add the new limit and after fields in the dump tests	2022-06-01 12:09:52 +02:00
Kerollmops	9d5cc88cd5	Implement the seek-based tasks list pagination	2022-06-01 12:09:52 +02:00
bors[bot]	d22f07f5b2	Merge #2448 2448: docs(security): Fix `Supported` r=MarinPostma a=ryanrussell Signed-off-by: Ryan Russell <git@ryanrussell.org> # Pull Request ## What does this PR do? - Fix typo in security `Suported` -> `Supported` ## PR checklist Please check if your PR fulfills the following requirements: - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Co-authored-by: Ryan Russell <git@ryanrussell.org>	2022-05-31 19:59:37 +00:00
bors[bot]	e81c7aa2e6	Merge #2423 2423: Paginate the index resource r=MarinPostma a=irevoire Fix #2373 Co-authored-by: Irevoire <tamo@meilisearch.com>	2022-05-31 19:25:25 +00:00
Ryan Russell	39db6ea42b	docs(security): Fix `Supported` Signed-off-by: Ryan Russell <git@ryanrussell.org>	2022-05-31 14:21:34 -05:00
bors[bot]	47007fa71b	Merge #2446 2446: rename Succeded to Succeeded r=irevoire a=MarinPostma this pr renames `TaskEvent::Succeded` to `TaskEvent::Succeeded` and apply the migration to the dumps Co-authored-by: ad hoc <postma.marin@protonmail.com>	2022-05-31 18:27:02 +00:00
Irevoire	627f13df85	feat(http): paginate the index resource Fix #2373	2022-05-31 18:11:45 +02:00
bors[bot]	97c14f6fcc	Merge #2427 2427: Update the documents resource r=MarinPostma a=irevoire - Return Documents API resources on `/documents` in an array in the results field. - Add limit, offset, and total in the response body. - Rename `attributesToRetrieve` into `fields` (only for the `/documents` endpoints, not for the `/search` ones). - The `displayedAttributes` settings do not impact anymore the displayed fields returned in the `/documents` endpoints. These settings only impact the `/search` endpoint. - make the `/document/:uid` route accept the `fields` query parameter Fix #2372 Co-authored-by: Irevoire <tamo@meilisearch.com>	2022-05-31 15:33:42 +00:00
ad hoc	446f1f31e0	rename Succeded to Succeeded	2022-05-31 17:22:37 +02:00
Irevoire	ddad6cc069	feat(http): update the documents resource - Return Documents API resources on `/documents` in an array in the the results field. - Add limit, offset and total in the response body. - Rename `attributesToRetrieve` into `fields` (only for the `/documents` endpoints, not for the `/search` ones). - The `displayedAttributes` settings does not impact anymore the displayed fields returned in the `/documents` endpoints. These settings only impacts the `/search` endpoint. Fix #2372	2022-05-31 16:40:40 +02:00
bors[bot]	ab39df9693	Merge #2399 2399: Update the tasks endpoints r=MarinPostma a=Kerollmops This PR wraps all the changes related to the `tasks` endpoints, it is related to https://github.com/meilisearch/meilisearch/issues/2377 but doesn't close it. I will create a new PR to work on [the seek-based pagination](https://github.com/meilisearch/specifications/pull/115). I wanted to do something cool with Github: being able to merge multiple PR in this one, to help review changes one by one, unfortunately, Github doesn't allow creating empty PRs. I also struggled with git itself when it comes to merging things in the right order, so I decided that I would add all of the changes in this single PR. I will list the changes and references to the specs here. - [x] Tasks statuses and types must be case insensitive - [x] Tasks statuses, types and indexUid must accept the `*` selector - [ ] Rename the `TaskDetails` struct fields ## Changes - [ ] Add seek-based pagination following [the spec](https://github.com/meilisearch/specifications/pull/115) - [x] Add filtering on the `/tasks` endpoint following [this spec](https://github.com/meilisearch/specifications/pull/116) - [x] Add filtering capabilities on `type`, `status` and `indexUid` for `GET` `task` lists endpoints. - [x] It is possible to specify several values for a filter using the `,` character. e.g. `?status=enqueued,processing` - [x] Between two different filters, an AND operation is applied. e.g. `?status=enqueued&type=indexCreation` is equivalent to `status=enqueued AND type = indexCreation` - [x] Remove `GET /indexes/:indexUid/tasks`. It can be replaced by `GET /tasks?indexUid=:indexUid` - [x] Remove `GET /indexes/:indexUid/tasks/:taskUid`. - [x] Rename `uid` to `taskUid` in the `202 - Accepted` task response return by every asynchronous tasks (ex: index creation, document addition...) - [x] Rename some task properties - [x] `documentPartial`-> `documentAdditionOrUpdate` - [x] `documentAddition`-> `documentAdditionOrUpdate` - [x] `clearAll` -> `documentDeletion` Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-05-31 09:40:40 +00:00
Kerollmops	1465b5e0ff	Refactorize the tasks filters by moving the match inside	2022-05-31 11:33:21 +02:00
Kerollmops	8800b348f0	Implement the StarOr on all the tasks filters	2022-05-31 11:33:21 +02:00
Kerollmops	082d6b89ff	Make the StarOrIndexUid Generic and call it StarOr	2022-05-31 11:33:21 +02:00
Kerollmops	b82c86c8f5	Allow users to filter indexUid with a *	2022-05-31 11:33:20 +02:00
Kerollmops	36d94257d8	Make clippy happy	2022-05-31 11:33:20 +02:00
Kerollmops	3f80468f18	Rename the Tasks Types	2022-05-31 11:33:20 +02:00
Kerollmops	8509243e68	Implement the status and type filtering on the tasks route	2022-05-31 11:33:20 +02:00
Kerollmops	3684c822f1	Add indexUid filtering on the /tasks route	2022-05-31 11:33:20 +02:00
Kerollmops	80f7d87356	Remove the /indexes/:indexUid/tasks/... routes	2022-05-31 11:33:20 +02:00
Kerollmops	d2f457a076	Rename the uid to taskUid in asynchronous response	2022-05-31 11:33:20 +02:00
Kerollmops	e5ef5a6f9c	Remove an unused updates.rs file	2022-05-31 11:33:19 +02:00
bors[bot]	5450fecaef	Merge #2444 2444: add boilerplate for dump v5 r=MarinPostma a=MarinPostma add the boilerplate files for dump v5 Co-authored-by: ad hoc <postma.marin@protonmail.com> Co-authored-by: ManyTheFish <many@meilisearch.com>	2022-05-31 08:56:52 +00:00
ManyTheFish	deba0cc096	Make v4::load_dump copy each part a the dump	2022-05-31 10:24:44 +02:00
ad hoc	26e7bdf702	add boilerplate for dump v5	2022-05-30 17:25:29 +02:00
bors[bot]	3441cc6c36	Merge #2410 2410: Make dump a task r=Kerollmops a=MarinPostma This PR transforms the dump task into a proper task. The `GET /dumps/:dump_uid` is removed. Some changes were made to make this work, and a bit a refactoring was necessary. - The `dump_actor` module has been renamed do `dumps` and moved to the root - There isn't a `DumpActor` anymore, and the dump process is handled by the `DumpHandler`. - The `TaskPerformer` is renamed to `BatchHandler` - The `BatchHandler` trait no longer has a `perform_job` method, but instead has a `accept` method returning whether a handler can proccess a batch - The scheduler now accept a list of `BatchHandler`, and iterates trhough them until it finds one to accept the current batch. - `Job` doesn't exist anymore, and everything in now inside of the `BatchContent` enum. - The `Vec<TaskId>` from `Batch` is replaced with a `BatchContent` enum which hints at the content. - The Scheduler is slightly modified to accept batch, and prioritize them before regular tasks. - The `TaskList` are not identified by a `String` representing the index uid anymore, but by a `TaskListIdentifier` which also works for dumps which are not targeting any specific indexes. - The `GET /dump/:dump_id` no longer exists - `DumpActorError` is renamed to `DumpError` close #2410 Co-authored-by: ad hoc <postma.marin@protonmail.com>	2022-05-30 14:09:43 +00:00
bors[bot]	c7711c7816	Merge #2429 2429: Send the analytics to `telemetry.meilisearch.com` instead of segment r=MarinPostma a=irevoire Fix #2425 Co-authored-by: Irevoire <tamo@meilisearch.com>	2022-05-30 13:18:55 +00:00
Irevoire	d47b997120	chore(analytics): update the url used to send our analytics	2022-05-30 15:13:10 +02:00
ad hoc	1e310ecc7d	fix typo in docstring Co-authored-by: Tamo <tamo@meilisearch.com>	2022-05-30 14:34:49 +02:00
ad hoc	4cb2c6ef1e	use map_or instead of map + unwrap_or	2022-05-30 12:30:15 +02:00
bors[bot]	582930dbbb	Merge #538 538: speedup exact words r=Kerollmops a=MarinPostma This PR make `exact_words` return an `Option` instead of an empty set, since set creation is costly, as noticed by `@kerollmops.` I was not convinces that this was the cause for all of the performance drop we measured, and then realized that methods that initialized it were called recursively which caused initialization times to add up. While the first fix solves the issue when not using exact words, using exact word remained way more expensive that it should be. To address this issue, the exact words are cached into the `Context`, so they are only initialized once. Co-authored-by: ad hoc <postma.marin@protonmail.com>	2022-05-30 08:20:34 +00:00
ad hoc	a9ef399a6b	processing::Nothing return BatchContent::Empty instead of panic	2022-05-26 12:04:27 +02:00
ad hoc	5a2972fc19	use TaskEvent method instead of variants in BatchHandler impl	2022-05-26 11:51:58 +02:00
0x0x1	ba51ca83ec	Update docker volume path Makes docker volume same as Dockerfile	2022-05-26 10:29:27 +08:00
ad hoc	1647ca3c1f	fix clipy warnings	2022-05-25 15:07:52 +02:00
ad hoc	74a1f88d88	add test for dump processing order	2022-05-25 14:57:36 +02:00
ad hoc	f58507379a	fix dump priority in scheduler	2022-05-25 14:50:14 +02:00
ad hoc	6b2016b350	remove typo in BatchContent variant	2022-05-25 14:39:07 +02:00
ad hoc	3015265bde	remove useless dump errors	2022-05-25 14:37:10 +02:00
ad hoc	49d8fadb52	test dump handler	2022-05-25 14:32:12 +02:00
ad hoc	127171c812	impl Default on Processing	2022-05-25 14:10:39 +02:00
bors[bot]	67b6f4340a	Merge #2422 2422: Update url of movies.json r=curquiza a=0x0x1 URL `https://bit.ly/2PAcw9l` is a notion site. Co-authored-by: 0x0x1 <101086451+0x0x1@users.noreply.github.com>	2022-05-25 11:21:56 +00:00
ad hoc	986a99296d	remove useless dump test	2022-05-25 11:25:11 +02:00
ad hoc	92d86ce6aa	add tests to IndexResolver BatchHandler	2022-05-25 11:13:36 +02:00
ad hoc	3c85b29865	add doc to BatchHandler	2022-05-25 11:13:35 +02:00
ad hoc	8349f38197	remove unused file	2022-05-25 11:13:35 +02:00
ad hoc	64654ef7c3	rename batch_handler to handler	2022-05-25 11:13:35 +02:00
ad hoc	0f9c134114	fix tests	2022-05-25 11:13:35 +02:00
ad hoc	7b47e4e87a	snapshot batch handler	2022-05-25 11:13:35 +02:00
ad hoc	8743d73973	move DumpHandler to own module	2022-05-25 11:13:35 +02:00
ad hoc	f0aceb4fba	remove unused files	2022-05-25 11:13:35 +02:00
ad hoc	61035a3ea4	create dump v5	2022-05-25 11:13:34 +02:00
ad hoc	4778884105	remove dump status route	2022-05-25 11:13:34 +02:00
ad hoc	57fde30b91	handle dump	2022-05-25 11:13:34 +02:00
ad hoc	56eb2907c9	dump indexes	2022-05-25 11:13:34 +02:00
ad hoc	414d0907ce	register dump handler	2022-05-25 11:13:34 +02:00
ad hoc	60a8249de6	add dump batch handler	2022-05-25 11:13:34 +02:00
ad hoc	46cdc17701	make scheduler accept multiple batch handlers	2022-05-25 11:13:34 +02:00
ad hoc	6a0231cb28	perform dump method	2022-05-25 11:13:33 +02:00
ad hoc	7fa3eb1003	register dump tasks	2022-05-25 11:13:33 +02:00
ad hoc	2f0625a984	register and insert dump task in scheduler	2022-05-25 11:13:33 +02:00
ad hoc	737b891a41	introduce Dump TaskListIdentifier variant	2022-05-25 11:13:33 +02:00
ad hoc	5a5066023b	introduce TaskListIdentifier	2022-05-25 11:13:33 +02:00
ad hoc	aa50acb031	make Task index_uid an option Not all task relate to an index. Tasks that don't have an index_uid set to None	2022-05-25 11:13:32 +02:00
bors[bot]	9935db86c7	Merge #2424 2424: Uncomment clippy from the ci check r=curquiza a=irevoire The issue has been fixed in the latest release of rust. See https://github.com/rust-lang/rust-clippy/issues/8662 Fix #2305 Co-authored-by: Irevoire <tamo@meilisearch.com>	2022-05-24 17:50:55 +00:00
bors[bot]	9f78e392b1	Merge #536 536: Improves ranking rules error message r=Kerollmops a=matthias-wright This PR improves the ranking rules error message to properly reflect the case sensitivity. The issue was highlighted in [meilisearch/issues/2407](https://github.com/meilisearch/meilisearch/issues/2407). Cheers! Co-authored-by: Matthias Wright <matthias.s.wright@gmail.com>	2022-05-24 16:43:52 +00:00
Irevoire	f65116b208	chore(ci): uncomment clippy from the ci check The issue has been fixed in the latest release of rust. See https://github.com/rust-lang/rust-clippy/issues/8662 Fix #2305	2022-05-24 15:03:11 +02:00
bors[bot]	341756a0eb	Merge #2357 2357: chore(dump): add dump tests r=Kerollmops a=irevoire Add tests on the import of dump v1, v2, v3 and v4. Since the dumps are slow to decompress, I made the `flate2` crate always compile in optimized. And since they're also slow to index, I also made the `milli` crate always compile in optimized. What do you think of this `@MarinPostma?` Should we keep milli unoptimized in case it could help us debug some things? 👀 Co-authored-by: Tamo <tamo@meilisearch.com>	2022-05-24 12:24:29 +00:00
Tamo	5f0e9b63d2	chore(dump): add tests	2022-05-24 14:21:56 +02:00
ad hoc	25fc576696	review changes	2022-05-24 14:15:33 +02:00
bors[bot]	ca9ba2d90c	Merge #2406 2406: chore(search): rename in the search endpoint r=irevoire a=irevoire Fix #2376 Co-authored-by: Irevoire <tamo@meilisearch.com>	2022-05-24 12:02:45 +00:00
ad hoc	69dc4de80f	change &Option<Set> to Option<&Set>	2022-05-24 12:14:55 +02:00
bors[bot]	2c248a68a4	Merge #2374 2374: feat(analytics): handle the `X-Meilisearch-Client` header r=Kerollmops a=irevoire Fix #2367 Co-authored-by: Tamo <tamo@meilisearch.com>	2022-05-24 09:49:16 +00:00
0x0x1	641ca5a857	Update url of movies.json	2022-05-24 17:34:34 +08:00
ad hoc	ac975cc747	cache context's exact words	2022-05-24 09:43:17 +02:00
ad hoc	8993fec8a3	return optional exact words	2022-05-24 09:15:49 +02:00
Tamo	6bf4db0bca	feat(analytics): handle the new x-meilisearch-client custom header for the analytics Fix #2367	2022-05-23 13:51:19 +02:00
Matthias Wright	754f48a4fb	Improves ranking rules error message	2022-05-20 21:25:43 +02:00
Irevoire	4e9accdeb7	chore(search): rename in the search endpoint Fix ##2376	2022-05-19 16:31:37 +02:00
bors[bot]	ae4e419db4	Merge #2408 2408: Upgrade milli v0.28 r=ManyTheFish a=ManyTheFish - Add smart crop - Add test on _matches_infos - Fix some test Co-authored-by: ManyTheFish <many@meilisearch.com>	2022-05-19 11:52:16 +00:00
ManyTheFish	50763aac82	Fix clippy	2022-05-19 11:23:22 +02:00
ManyTheFish	3517eae47f	Fix tests	2022-05-18 18:45:53 +02:00
ManyTheFish	0250ea9157	Intergrate smart crop in Meilisearch	2022-05-18 18:35:51 +02:00
Kerollmops	cd7c6e19ed	Reintroduce the max values by facet limit	2022-05-18 15:57:57 +02:00
bors[bot]	19dac01c5c	Merge #534 534: Bump milli version to v0.28.0 r=curquiza a=ManyTheFish Co-authored-by: ManyTheFish <many@meilisearch.com>	2022-05-18 09:04:46 +00:00
ManyTheFish	895f5d8a26	Bump milli version	2022-05-18 10:37:12 +02:00
bors[bot]	6d221058f1	Merge #2404 2404: Bring `release-v0.27.1` into `main` r=curquiza a=curquiza Following the v0.27.1 hotfixes Co-authored-by: ad hoc <postma.marin@protonmail.com> Co-authored-by: bors[bot] <26634292+bors[bot]@users.noreply.github.com> Co-authored-by: Tamo <tamo@meilisearch.com> Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-05-17 16:11:17 +00:00
bors[bot]	3389561f34	Merge #532 532: Add some implementation on MatchBounds r=Kerollmops a=ManyTheFish Theses Implementations are needed in meilisearch Co-authored-by: ManyTheFish <many@meilisearch.com>	2022-05-17 14:50:22 +00:00
ManyTheFish	137434a1c8	Add some implementation on MatchBounds	2022-05-17 15:57:09 +02:00
bors[bot]	08c6d50cd1	Merge #531 531: fix the mixed dataset geosearch indexing bug r=Kerollmops a=irevoire port #529 to main Co-authored-by: Tamo <tamo@meilisearch.com>	2022-05-16 16:06:36 +00:00
bors[bot]	cf3e574cb4	Merge #530 530: fix the searchable fields bug when a field is nested r=Kerollmops a=irevoire port #528 to main Co-authored-by: Tamo <tamo@meilisearch.com>	2022-05-16 15:52:30 +00:00
Tamo	0af399a6d7	fix the mixed dataset geosearch indexing bug	2022-05-16 17:37:45 +02:00
Tamo	f586028f9a	fix the searchable fields bug when a field is nested Update milli/src/index.rs Co-authored-by: Clément Renault <clement@meilisearch.com>	2022-05-16 17:24:36 +02:00
bors[bot]	e1e85267fd	Merge #526 526: remove useless comment r=irevoire a=MarinPostma Co-authored-by: ad hoc <postma.marin@protonmail.com>	2022-05-16 10:01:43 +00:00
bors[bot]	b9866d8df2	Merge #2339 2339: deny warnings in CI r=irevoire a=MarinPostma Add `RUSTFLAGS= -D warnings` to the CI so all warnings are treated as hard errors. Co-authored-by: ad hoc <postma.marin@protonmail.com>	2022-05-16 09:58:30 +00:00
bors[bot]	b9b9cba154	Merge #2383 2383: v0.27.0: bring `stable` into `main` r=Kerollmops a=curquiza Bring `stable` into `main` Co-authored-by: ad hoc <postma.marin@protonmail.com> Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com> Co-authored-by: bors[bot] <26634292+bors[bot]@users.noreply.github.com> Co-authored-by: ManyTheFish <many@meilisearch.com> Co-authored-by: Tamo <tamo@meilisearch.com> Co-authored-by: Paul Sanders <psanders1@gmail.com> Co-authored-by: Irevoire <tamo@meilisearch.com> Co-authored-by: Morgane Dubus <30866152+mdubus@users.noreply.github.com> Co-authored-by: Guillaume Mourier <guillaume@meilisearch.com>	2022-05-16 08:35:25 +00:00
bors[bot]	51809eb260	Merge #525 525: Simplify the error creation with thiserror r=irevoire a=irevoire I introduced [`thiserror`](https://docs.rs/thiserror/latest/thiserror/) to implements all the `Display` trait and most of the `impl From<xxx> for yyy` in way less lines. And then I introduced a cute macro to implements the `impl<X, Y, Z> From<X> for Z where Y: From<X>, Z: From<X>` more easily. Co-authored-by: Tamo <tamo@meilisearch.com>	2022-05-04 15:47:32 +00:00
Tamo	484a9ddb27	Simplify the error creation with thiserror and a smol friendly macro	2022-05-04 17:24:00 +02:00
bors[bot]	65e6aa0de2	Merge #523 523: Improve geosearch error messages r=irevoire a=irevoire Improve the geosearch error messages (#488). And try to parse the string as specified in https://github.com/meilisearch/meilisearch/issues/2354 Co-authored-by: Tamo <tamo@meilisearch.com>	2022-05-04 13:36:11 +00:00
ad hoc	348af6cfbf	deny-rust-warnings	2022-05-04 15:20:45 +02:00
bors[bot]	f3b9f7b867	Merge #527 527: Remove the wip section part of the contributing file r=curquiza a=Kerollmops Everything was good in the _Development Workflow_ section so I removed the _WIP Section_ part, now this PR fixes https://github.com/meilisearch/milli/issues/513. Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-05-04 13:11:30 +00:00
Kerollmops	48cdfddebf	Remove the wip section part of the contributing file	2022-05-04 14:44:51 +02:00
Tamo	c55368ddd4	apply code suggestion Co-authored-by: Kerollmops <kero@meilisearch.com>	2022-05-04 14:11:03 +02:00
bors[bot]	60ccb3fa4c	Merge #524 524: Add benchmark on nested fields r=irevoire a=irevoire fixes #500 Co-authored-by: Tamo <tamo@meilisearch.com>	2022-05-04 11:56:18 +00:00
ad hoc	5ad5d56f7e	remove useless comment	2022-05-04 10:43:54 +02:00
bors[bot]	0c2c8af44e	Merge #520 520: fix mistake in Settings initialization r=irevoire a=MarinPostma fix settings not being correctly initialized and add a test to make sure that they are in the future. fix https://github.com/meilisearch/meilisearch/issues/2358 Co-authored-by: ad hoc <postma.marin@protonmail.com>	2022-05-03 15:32:18 +00:00
bors[bot]	2fe9a02b1c	Merge #522 522: Do not generate keys that are too long for LMDB r=Kerollmops a=Kerollmops This PR fixes https://github.com/meilisearch/meilisearch/issues/2338 by making sure that we do not generate keys that are too long for LMDB especially when we are creating our prefix and proximity pairs keys. Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-05-03 11:54:10 +00:00
Kerollmops	211c8763b9	Make sure that we do not generate too long keys	2022-05-03 10:03:15 +02:00
Kerollmops	7e47031bdc	Add a test for long keys in LMDB	2022-05-03 10:03:13 +02:00
Tamo	f820c9804d	add one nested benchmark	2022-05-02 19:35:57 +02:00
Tamo	3cb1f6d0a1	improve geosearch error messages	2022-05-02 19:20:47 +02:00
ad hoc	1ee3d6ae33	fix mistake in Settings initialization	2022-04-29 16:24:25 +02:00
bors[bot]	312515dd6b	Merge #507 507: deny warnings in CI r=Kerollmops a=MarinPostma Add `RUSTFLAGS= -D warnings` to the CI so all warnings are treated as hard errors. Co-authored-by: ad hoc <postma.marin@protonmail.com>	2022-04-28 15:16:35 +00:00
ad hoc	3eb3f0269e	deny warnings in CI	2022-04-28 15:35:12 +02:00
bors[bot]	9db86aac51	Merge #518 518: Return facets even when there is no value associated to it r=Kerollmops a=Kerollmops This PR is related to https://github.com/meilisearch/meilisearch/issues/2352 and should fix the issue when Meilisearch is up-to-date with this PR. Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-04-28 09:04:36 +00:00
bors[bot]	2aae19dc52	Merge #517 517: Make nightly CI run every week r=Kerollmops a=curquiza Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-04-26 16:22:25 +00:00
Kerollmops	a4d343aade	Add a test to check for the returned facet distribution	2022-04-26 18:12:58 +02:00
bors[bot]	c2bd94c871	Merge #511 511: Update version in every workspace r=curquiza a=curquiza Checked with `@Kerollmops` - Update the version into every workspace (the current version is v0.27.0, but I forgot to update it for the previous release) - add `publish = false` except in `milli` workspace. Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-04-26 16:06:47 +00:00
Kerollmops	7d1c2d97bf	Return facets even when there is no values associated to it	2022-04-26 17:59:53 +02:00
bors[bot]	d388ea0f9d	Merge #506 506: fix cargo warnings r=Kerollmops a=MarinPostma fix cargo warnings Co-authored-by: ad hoc <postma.marin@protonmail.com>	2022-04-26 15:45:20 +00:00
Clémentine Urquizar	ec89030483	Update bors toml	2022-04-26 17:36:04 +02:00
ad hoc	5c29258e8e	fix cargo warnings	2022-04-26 17:33:11 +02:00
bors[bot]	2fdf520271	Merge #514 514: Stop flattening every field r=Kerollmops a=irevoire When we need to flatten a document: * The primary key contains a `.`. * Some fields need to be flattened Instead of flattening the whole object and thus creating a lot of allocations with the `serde_json_flatten_crate`, we instead generate a minimal sub-object containing only the fields that need to be flattened. That should create fewer allocations and thus index faster. --------- ``` group indexing_main_e1e362fa indexing_stop-flattening-every-field_40d1bd6b ----- ---------------------- --------------------------------------------- indexing/Indexing geo_point 1.99 23.7±0.23s ? ?/sec 1.00 11.9±0.21s ? ?/sec indexing/Indexing movies in three batches 1.00 18.2±0.24s ? ?/sec 1.01 18.3±0.29s ? ?/sec indexing/Indexing movies with default settings 1.00 17.5±0.09s ? ?/sec 1.01 17.7±0.26s ? ?/sec indexing/Indexing songs in three batches with default settings 1.00 64.8±0.47s ? ?/sec 1.00 65.1±0.49s ? ?/sec indexing/Indexing songs with default settings 1.00 54.9±0.99s ? ?/sec 1.01 55.7±1.34s ? ?/sec indexing/Indexing songs without any facets 1.00 50.6±0.62s ? ?/sec 1.01 50.9±1.05s ? ?/sec indexing/Indexing songs without faceted numbers 1.00 54.0±1.14s ? ?/sec 1.01 54.7±1.13s ? ?/sec indexing/Indexing wiki 1.00 996.2±8.54s ? ?/sec 1.02 1021.1±30.63s ? ?/sec indexing/Indexing wiki in three batches 1.00 1136.8±9.72s ? ?/sec 1.00 1138.6±6.59s ? ?/sec ``` So basically everything slowed down a liiiiiittle bit except the dataset with a nested field which got twice faster Co-authored-by: Tamo <tamo@meilisearch.com>	2022-04-26 11:50:33 +00:00
Tamo	f19d2dc548	Only flatten the required fields apply review comments Co-authored-by: Kerollmops <kero@meilisearch.com>	2022-04-26 12:33:46 +02:00
bors[bot]	5adeac8047	Merge #516 516: Fix the indexing fuzzer r=irevoire a=irevoire Co-authored-by: Tamo <tamo@meilisearch.com>	2022-04-26 08:35:03 +00:00
Clémentine Urquizar	7cb7643565	Make nightly CI run every week Update CI Fix CI	2022-04-25 18:52:27 +02:00
Clémentine Urquizar	d138b3c704	Update version	2022-04-25 18:43:46 +02:00
Tamo	fa6f495662	fix the indexing fuzzer	2022-04-25 18:32:06 +02:00
bors[bot]	8cc86d5a8d	Merge #515 515: Improve the README r=curquiza a=Kerollmops This PR closes #512 by adding more content to the README. We listed all of the subcrates of the repository, changed the descriptions of the subcrates, and added a simple example usage in the README. Co-authored-by: Kerollmops <clement@meilisearch.com> Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com>	2022-04-25 16:15:12 +00:00
Clémentine Urquizar - curqui	5e562ffecf	Update README.md	2022-04-25 18:14:43 +02:00
Clémentine Urquizar - curqui	2277172f9c	Update README.md	2022-04-25 18:14:39 +02:00
Clémentine Urquizar - curqui	2db3d60259	Update README.md	2022-04-25 18:14:35 +02:00
Kerollmops	7e19bf1c0e	Add an example usage of the library in the README	2022-04-25 17:25:46 +02:00
Kerollmops	fb192aaa9f	Update the list of milli's subcrates	2022-04-25 15:55:38 +02:00
bors[bot]	e1e362fa43	Merge #509 509: Remove pr_status from bors settings r=Kerollmops a=curquiza Because of multiple issue we had with bors. https://github.com/bors-ng/bors-ng/issues/1492 Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-04-25 11:45:37 +00:00
Clémentine Urquizar	08753d002a	Remove pr_status from bors settings	2022-04-25 13:39:45 +02:00
Clément Renault	8d15ae37a1	Merge pull request #503 from meilisearch/improve-flatten-fuzzer Improve the fuzzer of the flatten crate	2022-04-25 13:38:43 +02:00
Clément Renault	3e53791de3	Merge pull request #508 from meilisearch/contributing First version of new CONTRIBUTING.md	2022-04-25 13:36:41 +02:00
bors[bot]	8010eca9c7	Merge #505 505: normalize exact words r=curquiza a=MarinPostma Normalize the exact words, as specified in the specification. Co-authored-by: ad hoc <postma.marin@protonmail.com>	2022-04-25 09:35:32 +00:00
bors[bot]	c07f3b44b7	Merge #2347 2347: Change Nelson path r=curquiza a=curquiza Nelson is now on the Meilisearch orga side Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-04-21 17:50:46 +00:00
Clémentine Urquizar	dc0d4addd9	First version of new CONTRIBUTING.md	2022-04-21 19:02:22 +02:00
Clémentine Urquizar	38d681c230	Change Nelson path	2022-04-21 18:42:34 +02:00
Clémentine Urquizar - curqui	e85377e725	Merge pull request #2346 from meilisearch/revert-2345-bump-meilisearch-v9000.0.0 Revert "[TEST PURPOSE] Bump meilisearch to version 9000.0.0"	2022-04-21 16:38:48 +02:00
Clémentine Urquizar - curqui	6ff8bf823d	Revert "[TEST PURPOSE] Bump meilisearch to version 9000.0.0"	2022-04-21 16:36:56 +02:00
Clémentine Urquizar - curqui	4d25229df9	Merge pull request #2345 from meilisearch/bump-meilisearch-v9000.0.0 Bump meilisearch to version 9000.0.0	2022-04-21 16:28:46 +02:00
releasemops	f1cd6b6ee8	bump meilisearch to v9000.0.0	2022-04-21 14:26:40 +00:00
Clémentine Urquizar - curqui	63f75bd187	Merge pull request #2344 from meilisearch/revert-2340-bump-meilisearch-v8000.1.0 Revert "[TEST PURPOSE] Bump meilisearch to version 8000.1.0"	2022-04-21 16:24:57 +02:00
Clémentine Urquizar - curqui	acf3357cf3	Revert "[TEST PURPOSE] Bump meilisearch to version 8000.1.0"	2022-04-21 16:24:27 +02:00
Clément Renault	71414630fc	Merge pull request #504 from meilisearch/test-long-words Add a test to make sure that long words are handled	2022-04-21 16:06:13 +02:00
ad hoc	2e0089d5ff	normalize exact words	2022-04-21 15:38:40 +02:00
Clémentine Urquizar - curqui	202d6105b2	Merge pull request #2340 from meilisearch/bump-meilisearch-v8000.1.0 [TEST PURPOSE] Bump meilisearch to version 8000.1.0	2022-04-21 15:28:00 +02:00
releasemops	0714551101	bump meilisearch to v8000.1.0	2022-04-21 13:23:46 +00:00
ad hoc	3a2451fcba	add test normalize exact words	2022-04-21 13:52:09 +02:00
Clément Renault	eb5830aa40	Add a test to make sure that long words are handled	2022-04-21 13:45:28 +02:00
Tamo	d81a3f4a74	improve the fuzzer of the flatten crate	2022-04-20 16:11:23 +02:00
bors[bot]	c7d0097c97	Merge #498 498: Get rid of the threshold when comparing benchmarks r=curquiza a=irevoire It just hides things Co-authored-by: Tamo <tamo@meilisearch.com>	2022-04-19 14:04:11 +00:00
Tamo	152a10344c	Get rid of the threshold when comparing benchmarks It just hide things	2022-04-19 15:39:58 +02:00
bors[bot]	04eb32e539	Merge #499 499: fix min-word-len-for-typo not reset properly r=Kerollmops a=MarinPostma fix min word len for typo not resettign properly, as reported in https://github.com/meilisearch/meilisearch/issues/2330 Co-authored-by: ad hoc <postma.marin@protonmail.com>	2022-04-19 13:22:19 +00:00
ad hoc	8b14090927	fix min-word-len-for-typo not reset properly	2022-04-19 15:20:16 +02:00
bors[bot]	ea4bb9402f	Merge #483 483: Enhance matching words r=Kerollmops a=ManyTheFish # Summary Enhance milli word-matcher making it handle match computing and cropping. # Implementation ## Computing best matches for cropping Before we were considering that the first match of the attribute was the best one, this was accurate when only one word was searched but was missing the target when more than one word was searched. Now we are searching for the best matches interval to crop around, the chosen interval is the one: 1) that have the highest count of unique matches > for example, if we have a query `split the world`, then the interval `the split the split the` has 5 matches but only 2 unique matches (1 for `split` and 1 for `the`) where the interval `split of the world` has 3 matches and 3 unique matches. So the interval `split of the world` is considered better. 2) that have the minimum distance between matches > for example, if we have a query `split the world`, then the interval `split of the world` has a distance of 3 (2 between `split` and `the`, and 1 between `the` and `world`) where the interval `split the world` has a distance of 2. So the interval `split the world` is considered better. 3) that have the highest count of ordered matches > for example, if we have a query `split the world`, then the interval `the world split` has 2 ordered words where the interval `split the world` has 3. So the interval `split the world` is considered better. ## Cropping around the best matches interval Before we were cropping around the interval without checking the context. Now we are cropping around words in the same context as matching words. This means that we will keep words that are farther from the matching words but are in the same phrase, than words that are nearer but separated by a dot. > For instance, for the matching word `Split` the text: `Natalie risk her future. Split The World is a book written by Emily Henry. I never read it.` will be cropped like: `…. Split The World is a book written by Emily Henry. …` and not like: `Natalie risk her future. Split The World is a book …` Co-authored-by: ManyTheFish <many@meilisearch.com>	2022-04-19 11:42:32 +00:00
ManyTheFish	f1115e274f	Use Copy impl of FormatOption instead of clonning	2022-04-19 10:35:50 +02:00
Clémentine Urquizar - curqui	a68e3a79fb	Merge pull request #497 from meilisearch/v0.26.1 Update version for the next release (v0.26.1)	2022-04-14 11:53:31 +02:00
Clémentine Urquizar	8d630a6f62	Update version for the next release (v0.26.1)	2022-04-14 11:44:06 +02:00
Clémentine Urquizar - curqui	d362278a41	Merge pull request #494 from meilisearch/flatten-what-is-needed Only flatten the required objects	2022-04-14 11:43:28 +02:00
Tamo	00f78d6b5a	Apply code suggestions Co-authored-by: Clément Renault <clement@meilisearch.com>	2022-04-14 11:14:08 +02:00
Tamo	399fba16bb	only flatten an object if it's nested	2022-04-14 11:14:08 +02:00
Tamo	c2469b6765	create the json-depth-checker crate	2022-04-14 11:14:08 +02:00
bors[bot]	7791ef90e7	Merge #493 493: Use smartstring to store the external id in our hashmap r=Kerollmops a=irevoire We need to store all the external id (primary key) in a hashmap associated to their internal id. The smartstring remove heap allocation / memory usage and should improve the cache locality. I ran the benchmarks to measure the impact of this PR on the indexing time. I think we should merge it whatever happens thought because it'll decrease the memory consumption. --------- This improve really sliiiiiightly the performances but improve the memory usage thus it should be merged. ``` group indexing_main_6b073738 indexing_use-smartsring_3f343511 ----- ---------------------- -------------------------------- indexing/Indexing geo_point 1.02 25.2±0.20s ? ?/sec 1.00 24.8±0.13s ? ?/sec indexing/Indexing movies in three batches 1.00 18.2±0.10s ? ?/sec 1.00 18.2±0.23s ? ?/sec indexing/Indexing movies with default settings 1.00 17.5±0.09s ? ?/sec 1.01 17.7±0.11s ? ?/sec indexing/Indexing songs in three batches with default settings 1.00 68.3±1.01s ? ?/sec 1.00 68.0±0.95s ? ?/sec indexing/Indexing songs with default settings 1.00 63.2±0.78s ? ?/sec 1.00 63.0±0.58s ? ?/sec indexing/Indexing songs without any facets 1.02 59.6±1.00s ? ?/sec 1.00 58.5±1.03s ? ?/sec indexing/Indexing songs without faceted numbers 1.00 62.8±0.38s ? ?/sec 1.00 62.6±1.02s ? ?/sec indexing/Indexing wiki 1.01 1009.2±25.25s ? ?/sec 1.00 998.1±11.27s ? ?/sec indexing/Indexing wiki in three batches 1.01 1142.0±9.97s ? ?/sec 1.00 1134.4±11.21s ? ?/sec ``` Co-authored-by: Tamo <tamo@meilisearch.com>	2022-04-13 20:28:28 +00:00
Tamo	ee64f4a936	Use smartstring to store the external id in our hashmap We need to store all the external id (primary key) in a hashmap associated to their internal id during. The smartstring remove heap allocation / memory usage and should improve the cache locality.	2022-04-13 21:22:07 +02:00
bors[bot]	456887a54a	Merge #496 496: Improve the performances of the flattening subcrate r=irevoire a=Kerollmops This PR adds some benchmarks to the _flatten-serde-json_ crate, this crate is responsible for transforming the original documents into flat versions that the engine can understand. It can probably be speed-up and this is why I added benchmarks to it. I make some interesting performance improvements when I replaced the `json!` macro calls. ``` flatten/simple time: [452.44 ns 453.31 ns 454.18 ns] change: [-15.036% -14.751% -14.473%] (p = 0.00 < 0.05) Performance has improved. Found 2 outliers among 100 measurements (2.00%) 2 (2.00%) high mild Benchmarking flatten/complex: Collecting 100 samples in estimated 5.0007 s (4.9M i flatten/complex time: [1.0101 us 1.0131 us 1.0160 us] change: [-18.001% -17.775% -17.536%] (p = 0.00 < 0.05) Performance has improved. Found 6 outliers among 100 measurements (6.00%) 5 (5.00%) high mild 1 (1.00%) high severe ``` --- _I removed this particular commit from this PR._ The reason is that the two other commits were enough for this PR to give enough impact and be merged. We will continue to explore where we can get performances later. But when I changed the flattening function to accept an owned version of the objects, we lost a lot of performances. Yes, I rewrote the benchmarks (locally) to clone the input object (and measured both, previous and new versions, with the cloning benchmarks). Maybe cloning the benchmark inputs is not the right thing to do... ``` Benchmarking flatten/simple: Collecting 100 samples in estimated 5.0005 s (6.7M it flatten/simple time: [746.46 ns 749.59 ns 752.70 ns] change: [+40.082% +40.714% +41.347%] (p = 0.00 < 0.05) Performance has regressed. Benchmarking flatten/complex: Collecting 100 samples in estimated 5.0047 s (2.9M i flatten/complex time: [1.7311 us 1.7342 us 1.7368 us] change: [+40.976% +41.398% +41.807%] (p = 0.00 < 0.05) Performance has regressed. Found 1 outliers among 100 measurements (1.00%) 1 (1.00%) low mild ``` Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-04-13 11:14:29 +00:00
Kerollmops	b3cec1a383	Prefer using direct method calls instead of using the json macros	2022-04-13 13:12:57 +02:00
Kerollmops	436d2032c4	Add benchmarks to the flatten-serde-json subcrate	2022-04-13 13:12:57 +02:00
bors[bot]	3828635fb2	Merge #489 489: fix distinct count bug r=curquiza a=MarinPostma fix https://github.com/meilisearch/meilisearch/issues/2152 I think the issue was that we didn't take off the excluded candidates from the initial candidates when returning the candidates with the search result. Co-authored-by: ad hoc <postma.marin@protonmail.com>	2022-04-13 10:15:30 +00:00
ad hoc	dda28d7415	exclude excluded canditates from search result candidates	2022-04-13 12:10:35 +02:00
ad hoc	cd83014fff	add test for disctinct nb hits	2022-04-13 12:10:35 +02:00
ad hoc	bbb6728d2f	add distinct attributes to cli	2022-04-13 12:10:35 +02:00
bors[bot]	49fbbacafc	Merge #492 492: Add the new `Specify breaking` check to bors.toml r=curquiza a=curquiza Should prevent this problem: https://github.com/meilisearch/milli/pull/489#issuecomment-1094988060 Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com>	2022-04-13 08:59:40 +00:00
Clémentine Urquizar - curqui	7ad582f39f	Update bors.toml	2022-04-13 10:56:56 +02:00
Clémentine Urquizar - curqui	aa896f0e7a	Update bors.toml	2022-04-13 10:56:56 +02:00
Clémentine Urquizar - curqui	0261a0e3cf	Add the new `Specify breaking` check to bors.toml	2022-04-13 10:56:55 +02:00
ManyTheFish	5809d3ae0d	Add first benchmarks on formatting	2022-04-12 16:31:58 +02:00
ManyTheFish	827cedcd15	Add format option structure	2022-04-12 13:42:14 +02:00
ManyTheFish	011f8210ed	Make compute_matches more rust idiomatic	2022-04-12 10:19:02 +02:00
bors[bot]	6b0737384b	Merge #491 491: remove the unused key warning r=curquiza a=irevoire When I copy-pasted my flatten crate I forgot to remove the key used to publish the package and that throw a warning. Co-authored-by: Tamo <tamo@meilisearch.com>	2022-04-11 16:55:25 +00:00
Tamo	e153418b8a	remove the unused key warning	2022-04-11 14:52:41 +02:00
bors[bot]	c8306616e0	Merge #490 490: Enforce labelling for the PRs r=curquiza a=curquiza - Enforce one of the following labels to make the CI pass: `no breaking`, `DB breaking`, `API breaking` (milli API, not the Meilisearch API of course), or `skip changelog`. This new CI is now `Required` in the GitHub settings for merging a PR. - Adapt the release drafter to these new labels - rename `skip-changelog` into `skip changelog` according to the new label name Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-04-11 08:24:23 +00:00
Clémentine Urquizar	9383629d13	Enforce labelling for the PRs	2022-04-09 23:47:06 +02:00
ManyTheFish	a16de5de84	Symplify format and remove intermediate function	2022-04-08 11:20:41 +02:00
ManyTheFish	a769e09dfa	Make token_crop_bounds more rust idiomatic	2022-04-07 20:15:14 +02:00
bors[bot]	9ac2fd1c37	Merge #487 487: Update version (v0.26.0) r=Kerollmops a=curquiza breaking because of #458 Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-04-07 17:10:24 +00:00
bors[bot]	80ae020bee	Merge #458 458: Nested fields r=Kerollmops a=irevoire For the following document: ```json { "id": 1, "person": { "name": "tamo", "age": 25, } } ``` Suppose the user sets `person` as a filterable attribute. We need to store `person` in the filterable _obviously_. But we also need to keep track of `person.name` and `person.age` somewhere. That’s where I changed a little bit the logic of the engine. Currently, we have a function called `faceted_field` that returns the union of the filterable and sortable. I renamed this function in `user_defined_faceted_field`. And now, when we finish indexing documents, we look at all the fields and see if they « match » a `user_defined_faceted_field`. So in our case: - does `id` match `person`: 🔴 - does `person.name` match `person`: 🟢 - does `person.age` match `person`: 🟢 And thus, we insert in the database the following faceted fields: `person, person.name, person.age`. The good thing about that solution is that we generate everything during the indexing phase, and then during the search, we can access our field without recomputing too much globbing. ----- Now the bad thing is that I had to create a new db. And if that was only one db, that would be ok, but actually, I need to do the same for the: - Displayed attributes - Attributes to retrieve - Attributes to highlight - Attribute to crop `@Kerollmops` Do you think there is a better way to do it? Apart from all the code, can we have a problem because we have too many dbs? Co-authored-by: Irevoire <tamo@meilisearch.com> Co-authored-by: Tamo <tamo@meilisearch.com>	2022-04-07 16:26:09 +00:00
Tamo	bab898ce86	move the flatten-serde-json crate inside of milli	2022-04-07 18:20:44 +02:00
ManyTheFish	c8ed1675a7	Add some documentation	2022-04-07 17:32:13 +02:00
ManyTheFish	b1905dfa24	Make split_best_frequency returns references instead of owned data	2022-04-07 17:05:44 +02:00
Tamo	ab458d8840	fix tests after rebase	2022-04-07 17:00:00 +02:00
Irevoire	4f3ce6d9cd	nested fields	2022-04-07 16:58:46 +02:00
Clémentine Urquizar	ee1d627803	Update version (v0.26.0)	2022-04-07 15:56:10 +02:00
bors[bot]	4ae7aea3b2	Merge #486 486: Update version (v0.25.0) r=curquiza a=curquiza v0.25.0 will be released once #478 is merged Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-04-06 11:40:41 +00:00
bors[bot]	aadb0c58c9	Merge #478 478: Disable typo on attribute r=Kerollmops a=MarinPostma disable typo on attributes Co-authored-by: ad hoc <postma.marin@protonmail.com>	2022-04-05 23:45:35 +00:00
ad hoc	86249e2ae4	add missing \t in cli update display Co-authored-by: Clément Renault <clement@meilisearch.com>	2022-04-05 21:35:06 +02:00
ad hoc	b799f3326b	rename merge_nothing to merge_ignore_values	2022-04-05 18:44:35 +02:00
ManyTheFish	fa7d3a37c0	Make some cleaning and add comments	2022-04-05 17:48:56 +02:00
ManyTheFish	3bb1e35ada	Fix match count	2022-04-05 17:48:45 +02:00
ManyTheFish	56e0edd621	Put crop markers direclty around words	2022-04-05 17:41:32 +02:00
ManyTheFish	a93cd8c61c	Fix prefix highlight with special chars	2022-04-05 17:41:32 +02:00
ManyTheFish	b3f0f39106	Make some cleaning	2022-04-05 17:41:32 +02:00
ManyTheFish	6dc345bc53	Test and Fix prefix highlight	2022-04-05 17:41:32 +02:00
ManyTheFish	bd30ee97b8	Keep separators at start of the croped string	2022-04-05 17:41:32 +02:00
ManyTheFish	29c5f76d7f	Use new matcher in http-ui	2022-04-05 17:41:32 +02:00
ManyTheFish	734d0899d3	Publish Matcher	2022-04-05 17:41:32 +02:00
ManyTheFish	4428cb5909	Add some tests and fix some corner cases	2022-04-05 17:41:32 +02:00
ManyTheFish	844f546a8b	Add matches algorithm V1	2022-04-05 17:41:32 +02:00
ManyTheFish	3be1790803	Add crop algorithm with naive match algorithm	2022-04-05 17:41:32 +02:00
ManyTheFish	d96e72e5dc	Create formater with some tests	2022-04-05 17:41:32 +02:00
ad hoc	201fea0fda	limit extract_word_docids memory usage	2022-04-05 14:14:15 +02:00
ad hoc	5cfd3d8407	add exact attributes documentation	2022-04-05 14:10:22 +02:00
Clémentine Urquizar	9eec44dd98	Update version (v0.25.0)	2022-04-05 12:06:42 +02:00
ad hoc	b85cd4983e	remove field_id_from_position	2022-04-05 09:50:34 +02:00
ad hoc	dac81b2d44	add missing \n in cli settings	2022-04-05 09:48:56 +02:00
ad hoc	ab185a59b5	fix infos	2022-04-05 09:46:56 +02:00
ad hoc	59e41d98e3	add comments to integration test	2022-04-04 21:17:06 +02:00
ad hoc	1810927dbd	rephrase exact_attributes doc	2022-04-04 21:04:49 +02:00
ad hoc	b7694c34f5	remove println	2022-04-04 21:00:07 +02:00
ad hoc	6cabd47c32	fix typo in comment	2022-04-04 20:59:20 +02:00
ad hoc	9963f11172	fix infos crate compilation issue	2022-04-04 20:54:03 +02:00
ad hoc	c8d3a09af8	add integration test for disabel typo on attributes	2022-04-04 20:54:03 +02:00
ad hoc	bfd81ce050	add exact atttributes to cli settings	2022-04-04 20:54:03 +02:00
ad hoc	6b2c2509b2	fix bug in exact search	2022-04-04 20:54:03 +02:00
ad hoc	56b4f5dce2	add exact prefix to query_docids	2022-04-04 20:54:03 +02:00
ad hoc	21ae4143b1	add exact_word_prefix to Context	2022-04-04 20:54:03 +02:00
ad hoc	e8f06f6c06	extract exact_word_prefix_docids	2022-04-04 20:54:03 +02:00
ad hoc	6dd2e4ffbd	introduce exact_word_prefix database in index	2022-04-04 20:54:03 +02:00
ad hoc	ba0bb29cd8	refactor WordPrefixDocids to take dbs instead of indexes	2022-04-04 20:54:02 +02:00
ad hoc	c4c6e35352	query exact_word_docids in resolve_query_tree	2022-04-04 20:54:02 +02:00
ad hoc	8d46a5b0b5	extract exact word docids	2022-04-04 20:54:02 +02:00
ad hoc	5451c64d5d	increase criteria asc desc test map size	2022-04-04 20:54:02 +02:00
ad hoc	0a77be4ec0	introduce exact_word_docids db	2022-04-04 20:54:02 +02:00
ad hoc	5f9f82757d	refactor spawn_extraction_task	2022-04-04 20:54:02 +02:00
ad hoc	f82d4b36eb	introduce exact attribute setting	2022-04-04 20:54:02 +02:00
ad hoc	c882d8daf0	add test for exact words	2022-04-04 20:54:01 +02:00
ad hoc	7e9d56a9e7	disable typos on exact words	2022-04-04 20:54:01 +02:00
bors[bot]	900825bac0	Merge #474 474: Disable typos on exact word r=MarinPostma a=MarinPostma This PR introduces the `exact_word` setting to disable typo tolerance on custom words. If a user query contains a word from `exact_words`, no typo derivation will be made for that particular word. I have chosen to store the words in a FST, to save on deserialization, and allow for fast lookups. I had some trouble with the `serde` module, and had to rename it `serde_impl`. ## steps: - [x] introduce new settings to register words to disable typos on - [x] in `typos`, return exact match is the current word is part of the word to disable typos for. - [x] update `Context` to return the exact words dictionary. - [x] merge #473 Co-authored-by: ad hoc <postma.marin@protonmail.com>	2022-04-04 18:39:43 +00:00
ad hoc	3e67d8818c	fix typo in test comment	2022-04-04 20:34:23 +02:00
ad hoc	284d8a24e0	add intergration test for disabled typon on word	2022-04-04 20:15:51 +02:00
ad hoc	30a2711bac	rename serde module to serde_impl module needed because of issues with rustfmt	2022-04-04 20:10:55 +02:00
ad hoc	0fd55db21c	fmt	2022-04-04 20:10:55 +02:00
ad hoc	559e46be5e	fix bad rebase bug	2022-04-04 20:10:55 +02:00
ad hoc	8b1e5d9c6d	add test for exact words	2022-04-04 20:10:55 +02:00
ad hoc	774fa8f065	disable typos on exact words	2022-04-04 20:10:55 +02:00
ad hoc	9bbffb8fee	add exact words setting	2022-04-04 20:10:54 +02:00
bors[bot]	48a5ce7434	Merge #473 473: set minimum word len for typos r=MarinPostma a=MarinPostma this PR allows the configuration on the minimum word length for typos. The default values are the same as previously. ## steps - [x] introduce settings for the minimum word length for 1 and 2 typos - [x] update the settings update flow to set this setting - [x] create a structure `TypoConfig` to configure typo tolerance in the query builder - [x] in `typo`, use the configuration to create the appropriate query tree node. - [x] extend `Context` to return the setting for minimum word length for typos - [x] return correct error message for wrong settings. - [x] merge #469 Co-authored-by: ad hoc <postma.marin@protonmail.com>	2022-04-04 17:53:14 +00:00
bors[bot]	6bf9824fec	Merge #485 485: fix bug on 2 typos derivation r=Kerollmops a=MarinPostma I found a bug while working on #473. This pr fixes it and add the missing tests on word derivations. Co-authored-by: ad hoc <postma.marin@protonmail.com>	2022-04-04 17:17:53 +00:00
ad hoc	853b4a520f	fmt	2022-04-04 10:41:46 +02:00
ad hoc	2cb71dff4a	add typo integration tests	2022-04-04 10:41:46 +02:00
ad hoc	1941072bb2	implement Copy on Setting	2022-04-04 10:41:46 +02:00
ad hoc	fdaf45aab2	replace hardcoded value with constant in TestContext	2022-04-04 10:41:46 +02:00
ad hoc	950a740bd4	refactor typos for readability	2022-04-04 10:41:46 +02:00
ad hoc	66020cd923	rename min_word_len* to use plain letter numbers	2022-04-04 10:41:46 +02:00
ad hoc	4c4b336ecb	rename min word len for typo error	2022-04-01 11:17:03 +02:00
ad hoc	286dd7b2e4	rename min_word_len_2_typo	2022-04-01 11:17:03 +02:00
ad hoc	55af85db3c	add tests for min_word_len_for_typo	2022-04-01 11:17:02 +02:00
ad hoc	9102de5500	fix error message	2022-04-01 11:17:02 +02:00
ad hoc	a1a3a49bc9	dynamic minimum word len for typos in query tree builder	2022-04-01 11:17:02 +02:00
ad hoc	5a24e60572	introduce word len for typo setting	2022-04-01 11:17:02 +02:00
ad hoc	9fe40df960	add word derivations tests	2022-04-01 11:05:18 +02:00
ad hoc	d5ddc6b080	fix 2 typos word derivation bug	2022-04-01 10:51:22 +02:00
bors[bot]	d2d930dd3f	Merge #469 469: add authorize typo setting r=Kerollmops a=MarinPostma This PR adds support for an authorize typo settings. This makes is possible to disable typos for a whole index. Typos are enabled by default. Co-authored-by: ad hoc <postma.marin@protonmail.com>	2022-03-31 15:18:08 +00:00
ad hoc	3e34981d9b	add test for authorize_typos in update	2022-03-31 14:12:00 +02:00
ad hoc	6ef3bb9d83	fmt	2022-03-31 14:06:23 +02:00
ad hoc	f782fe2062	add authorize_typo_test	2022-03-31 10:08:39 +02:00
ad hoc	c4653347fd	add authorize typo setting	2022-03-31 10:05:44 +02:00
bors[bot]	d8dd357326	Merge #480 480: Increase benchmarks (push) CI timeout r=Kerollmops a=Kerollmops This PR fixes the fact that the benchmarks CI on push were [canceled by GitHub](https://github.com/meilisearch/milli/actions/runs/2028844132) because they reached the default timeout of 6h. This PR changes the timeout to 72h, the same setting as the manually triggered benchmark one. Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-03-29 18:13:31 +00:00
Kerollmops	6a77c81a28	Increase benchmarks (push) CI timeout	2022-03-29 09:45:36 -07:00
bors[bot]	e10c26e70d	Merge #479 479: Update version (v0.24.1) r=Kerollmops a=curquiza From v0.23.1 to v0.24.1 since we had an issue with the versionning for the previous release Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-03-24 20:12:37 +00:00
Clémentine Urquizar	ddf78a735b	Update version (v0.24.1)	2022-03-24 16:39:45 +01:00
bors[bot]	2c7cafbf20	Merge #475 475: Bump tokenizer r=Kerollmops a=irevoire This PR bump the tokenizer in v0.2.9 which fixes an issue we had with lindera where reqwest was used with openssl (which was breaking our benchmarks). Co-authored-by: Irevoire <tamo@meilisearch.com>	2022-03-23 13:26:44 +00:00
Irevoire	86dd88698d	bump tokenizer	2022-03-23 14:25:58 +01:00
bors[bot]	b82f46e862	Merge #476 476: Rollback meilisearch-tokenizer version r=Kerollmops a=irevoire Lindera often fails to download some data from google drive we can’t compile consistently meilisearch / milli. We can’t bump to the latest version (that moved out of google drive) either because lindera uses reqwest with openssl with no way of configuring it our benchmarks were not able to run. The latter issue should be fixed by https://github.com/lindera-morphology/lindera/pull/164. Co-authored-by: Irevoire <tamo@meilisearch.com>	2022-03-22 14:02:00 +00:00
Irevoire	5dc464b9a7	rollback meilisearch-tokenizer version	2022-03-21 17:29:10 +01:00
bors[bot]	90276d9a2d	Merge #472 472: Remove useless variables in proximity r=Kerollmops a=ManyTheFish Was passing by plane sweep algorithm to find some inspiration, and I discover that we have useless variables that were not detected because of the recursive function. Co-authored-by: ManyTheFish <many@meilisearch.com>	2022-03-16 15:33:11 +00:00
ManyTheFish	49d59d88c2	Remove useless variables in proximity	2022-03-16 16:12:52 +01:00
bors[bot]	5863afa1a5	Merge #468 468: Add a new error message when the filterableAttributes are empty r=Kerollmops a=brunoocasali Fixes https://github.com/meilisearch/meilisearch/issues/2140 Is there a good way to reduce de duplication here? Maybe adding a shared function? I don't know the best and idiomatic way to do that, I appreciate any tip! Another doubt is related to the duplication of the calling: ```rs // filter.rs:373 FilterError::AttributeNotFilterable { attribute, filterable: filterable_fields.into_iter().collect::<Vec<_>>().join(" "), }, ``` and ```rs // filter.rs:424 return Err(point[0].as_external_error(FilterError::AttributeNotFilterable { attribute: "_geo", filterable: filterable_fields.into_iter().collect::<Vec<_>>().join(" "), }))?; ``` I think we could make the `filterable_fields.into_iter().collect::<Vec<_>>().join(" ")` directly into the error handling like the sortable error. I made it into the last commit, if this is something to avoid, let me know and I can remove it :) Co-authored-by: Bruno Casali <brunoocasali@gmail.com>	2022-03-16 15:02:19 +00:00
Bruno Casali	adc71742c8	Move string concat to the struct instead of in the calling	2022-03-16 10:26:12 -03:00
bors[bot]	cb6b6915a4	Merge #470 470: Set the cargo crate resolver to v2 r=Kerollmops a=MarinPostma This PR updates the workspace resolver to v2. This should fix [the benchmarks](https://github.com/meilisearch/milli/runs/5558347765?check_suite_focus=true#step:8:184). Co-authored-by: ad hoc <postma.marin@protonmail.com>	2022-03-16 10:55:22 +00:00
ad hoc	2a31cd13c9	set resolver to v2	2022-03-16 11:47:27 +01:00
Bruno Casali	4822fe1beb	Add a better error message when the filterable attrs are empty Fixes https://github.com/meilisearch/meilisearch/issues/2140	2022-03-15 18:13:59 -03:00
bors[bot]	f04ab67083	Merge #466 466: Bump version to 0.23.1 r=curquiza a=Kerollmops This PR bumps the crate versions to 0.23.1. Nothing seems to be breaking in the next release. Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-03-15 17:19:05 +00:00
bors[bot]	ad4c982c68	Merge #439 439: Optimize typo criterion r=Kerollmops a=MarinPostma This pr implements a couple of optimization for the typo criterion: - clamp max typo on concatenated query words to 1: By considering that a concatenated query word is a typo, we clamp the max number of typos allowed o it to 1. This is useful because we noticed that concatenated query words often introduced words with 2 typos in queries that otherwise didn't allow for 2 typo words. - Make typos on the first letter count for 2. This change is a big performance gain: by considering the typos on the first letter to count as 2 typos, we drastically restrict the search space for 1 typo, and if we reach 2 typos, the search space is reduced as well, as we only consider: (2 typos ∩ correct first letter) ∪ (wrong first letter ∩ 1 typo) instead of 2 typos anywhere in the word. ## benches ``` group main typo ----- ---- ---- smol-songs.csv: asc + default/Notstandskomitee 2.51 5.8±0.01ms ? ?/sec 1.00 2.3±0.01ms ? ?/sec smol-songs.csv: asc + default/charles 2.48 3.0±0.01ms ? ?/sec 1.00 1190.9±1.29µs ? ?/sec smol-songs.csv: asc + default/charles mingus 5.56 10.8±0.01ms ? ?/sec 1.00 1935.3±1.00µs ? ?/sec smol-songs.csv: asc + default/david 1.65 3.9±0.00ms ? ?/sec 1.00 2.4±0.01ms ? ?/sec smol-songs.csv: asc + default/david bowie 3.34 12.5±0.02ms ? ?/sec 1.00 3.7±0.00ms ? ?/sec smol-songs.csv: asc + default/john 1.00 1849.7±3.74µs ? ?/sec 1.01 1875.1±4.65µs ? ?/sec smol-songs.csv: asc + default/marcus miller 4.32 15.7±0.01ms ? ?/sec 1.00 3.6±0.01ms ? ?/sec smol-songs.csv: asc + default/michael jackson 3.31 12.5±0.01ms ? ?/sec 1.00 3.8±0.00ms ? ?/sec smol-songs.csv: asc + default/tamo 1.05 565.4±0.86µs ? ?/sec 1.00 539.3±1.22µs ? ?/sec smol-songs.csv: asc + default/thelonious monk 3.49 11.5±0.01ms ? ?/sec 1.00 3.3±0.00ms ? ?/sec smol-songs.csv: asc/Notstandskomitee 2.59 5.6±0.02ms ? ?/sec 1.00 2.2±0.01ms ? ?/sec smol-songs.csv: asc/charles 6.05 2.1±0.00ms ? ?/sec 1.00 347.8±0.60µs ? ?/sec smol-songs.csv: asc/charles mingus 14.46 9.4±0.01ms ? ?/sec 1.00 649.2±0.97µs ? ?/sec smol-songs.csv: asc/david 3.87 2.4±0.00ms ? ?/sec 1.00 618.2±0.69µs ? ?/sec smol-songs.csv: asc/david bowie 10.14 9.8±0.01ms ? ?/sec 1.00 970.8±1.55µs ? ?/sec smol-songs.csv: asc/john 1.00 546.5±1.10µs ? ?/sec 1.00 547.1±2.11µs ? ?/sec smol-songs.csv: asc/marcus miller 11.45 10.4±0.06ms ? ?/sec 1.00 907.9±1.37µs ? ?/sec smol-songs.csv: asc/michael jackson 10.56 9.7±0.01ms ? ?/sec 1.00 919.6±1.03µs ? ?/sec smol-songs.csv: asc/tamo 1.03 43.3±0.18µs ? ?/sec 1.00 42.2±0.23µs ? ?/sec smol-songs.csv: asc/thelonious monk 4.16 10.7±0.02ms ? ?/sec 1.00 2.6±0.00ms ? ?/sec smol-songs.csv: basic filter: <=/Notstandskomitee 1.00 95.7±0.20µs ? ?/sec 1.15 109.6±10.40µs ? ?/sec smol-songs.csv: basic filter: <=/charles 1.00 27.8±0.15µs ? ?/sec 1.01 27.9±0.18µs ? ?/sec smol-songs.csv: basic filter: <=/charles mingus 1.72 119.2±0.67µs ? ?/sec 1.00 69.1±0.13µs ? ?/sec smol-songs.csv: basic filter: <=/david 1.00 22.3±0.33µs ? ?/sec 1.05 23.4±0.19µs ? ?/sec smol-songs.csv: basic filter: <=/david bowie 1.59 86.9±0.79µs ? ?/sec 1.00 54.5±0.31µs ? ?/sec smol-songs.csv: basic filter: <=/john 1.00 17.9±0.06µs ? ?/sec 1.06 18.9±0.15µs ? ?/sec smol-songs.csv: basic filter: <=/marcus miller 1.65 102.7±1.63µs ? ?/sec 1.00 62.3±0.18µs ? ?/sec smol-songs.csv: basic filter: <=/michael jackson 1.76 128.2±1.85µs ? ?/sec 1.00 72.9±0.19µs ? ?/sec smol-songs.csv: basic filter: <=/tamo 1.00 17.9±0.13µs ? ?/sec 1.05 18.7±0.20µs ? ?/sec smol-songs.csv: basic filter: <=/thelonious monk 1.53 157.5±2.38µs ? ?/sec 1.00 102.8±0.88µs ? ?/sec smol-songs.csv: basic filter: TO/Notstandskomitee 1.00 100.9±4.36µs ? ?/sec 1.04 105.0±8.25µs ? ?/sec smol-songs.csv: basic filter: TO/charles 1.00 28.4±0.36µs ? ?/sec 1.03 29.4±0.33µs ? ?/sec smol-songs.csv: basic filter: TO/charles mingus 1.71 118.1±1.08µs ? ?/sec 1.00 68.9±0.26µs ? ?/sec smol-songs.csv: basic filter: TO/david 1.00 24.0±0.26µs ? ?/sec 1.03 24.6±0.43µs ? ?/sec smol-songs.csv: basic filter: TO/david bowie 1.72 95.2±0.30µs ? ?/sec 1.00 55.2±0.14µs ? ?/sec smol-songs.csv: basic filter: TO/john 1.00 18.8±0.09µs ? ?/sec 1.06 19.8±0.17µs ? ?/sec smol-songs.csv: basic filter: TO/marcus miller 1.61 102.4±1.65µs ? ?/sec 1.00 63.4±0.24µs ? ?/sec smol-songs.csv: basic filter: TO/michael jackson 1.77 132.1±1.41µs ? ?/sec 1.00 74.5±0.59µs ? ?/sec smol-songs.csv: basic filter: TO/tamo 1.00 18.2±0.14µs ? ?/sec 1.05 19.2±0.46µs ? ?/sec smol-songs.csv: basic filter: TO/thelonious monk 1.49 150.8±1.92µs ? ?/sec 1.00 101.3±0.44µs ? ?/sec smol-songs.csv: basic placeholder/ 1.00 27.3±0.07µs ? ?/sec 1.03 28.0±0.05µs ? ?/sec smol-songs.csv: basic with quote/"Notstandskomitee" 1.00 122.4±0.17µs ? ?/sec 1.03 125.6±0.16µs ? ?/sec smol-songs.csv: basic with quote/"charles" 1.00 88.8±0.30µs ? ?/sec 1.00 88.4±0.15µs ? ?/sec smol-songs.csv: basic with quote/"charles" "mingus" 1.00 685.2±0.74µs ? ?/sec 1.01 689.4±6.07µs ? ?/sec smol-songs.csv: basic with quote/"david" 1.00 161.6±0.42µs ? ?/sec 1.01 162.6±0.17µs ? ?/sec smol-songs.csv: basic with quote/"david" "bowie" 1.00 731.7±0.73µs ? ?/sec 1.02 743.1±0.77µs ? ?/sec smol-songs.csv: basic with quote/"john" 1.00 267.1±0.33µs ? ?/sec 1.01 270.9±0.33µs ? ?/sec smol-songs.csv: basic with quote/"marcus" "miller" 1.00 138.7±0.31µs ? ?/sec 1.02 140.9±0.13µs ? ?/sec smol-songs.csv: basic with quote/"michael" "jackson" 1.01 841.4±0.72µs ? ?/sec 1.00 833.8±0.92µs ? ?/sec smol-songs.csv: basic with quote/"tamo" 1.01 189.2±0.26µs ? ?/sec 1.00 188.2±0.71µs ? ?/sec smol-songs.csv: basic with quote/"thelonious" "monk" 1.00 1100.5±1.36µs ? ?/sec 1.01 1111.7±2.17µs ? ?/sec smol-songs.csv: basic without quote/Notstandskomitee 3.40 7.9±0.02ms ? ?/sec 1.00 2.3±0.02ms ? ?/sec smol-songs.csv: basic without quote/charles 2.57 494.4±0.89µs ? ?/sec 1.00 192.5±0.18µs ? ?/sec smol-songs.csv: basic without quote/charles mingus 1.29 2.8±0.02ms ? ?/sec 1.00 2.1±0.01ms ? ?/sec smol-songs.csv: basic without quote/david 1.95 623.8±0.90µs ? ?/sec 1.00 319.2±1.22µs ? ?/sec smol-songs.csv: basic without quote/david bowie 1.12 5.9±0.00ms ? ?/sec 1.00 5.2±0.00ms ? ?/sec smol-songs.csv: basic without quote/john 1.24 1340.9±2.25µs ? ?/sec 1.00 1084.7±7.76µs ? ?/sec smol-songs.csv: basic without quote/marcus miller 7.97 14.6±0.01ms ? ?/sec 1.00 1826.0±6.84µs ? ?/sec smol-songs.csv: basic without quote/michael jackson 1.19 3.9±0.00ms ? ?/sec 1.00 3.3±0.00ms ? ?/sec smol-songs.csv: basic without quote/tamo 1.65 737.7±3.58µs ? ?/sec 1.00 446.7±0.51µs ? ?/sec smol-songs.csv: basic without quote/thelonious monk 1.16 4.5±0.02ms ? ?/sec 1.00 3.9±0.04ms ? ?/sec smol-songs.csv: big filter/Notstandskomitee 3.27 7.6±0.02ms ? ?/sec 1.00 2.3±0.01ms ? ?/sec smol-songs.csv: big filter/charles 8.26 1957.5±1.37µs ? ?/sec 1.00 236.8±0.34µs ? ?/sec smol-songs.csv: big filter/charles mingus 18.49 11.2±0.06ms ? ?/sec 1.00 607.7±3.03µs ? ?/sec smol-songs.csv: big filter/david 3.78 2.4±0.00ms ? ?/sec 1.00 622.8±0.80µs ? ?/sec smol-songs.csv: big filter/david bowie 9.00 12.0±0.01ms ? ?/sec 1.00 1336.0±3.17µs ? ?/sec smol-songs.csv: big filter/john 1.00 554.2±0.95µs ? ?/sec 1.01 560.4±0.79µs ? ?/sec smol-songs.csv: big filter/marcus miller 18.09 12.0±0.01ms ? ?/sec 1.00 664.7±0.60µs ? ?/sec smol-songs.csv: big filter/michael jackson 8.43 12.0±0.01ms ? ?/sec 1.00 1421.6±1.37µs ? ?/sec smol-songs.csv: big filter/tamo 1.00 86.3±0.14µs ? ?/sec 1.01 87.3±0.21µs ? ?/sec smol-songs.csv: big filter/thelonious monk 5.55 14.3±0.02ms ? ?/sec 1.00 2.6±0.01ms ? ?/sec smol-songs.csv: desc + default/Notstandskomitee 2.52 5.8±0.01ms ? ?/sec 1.00 2.3±0.01ms ? ?/sec smol-songs.csv: desc + default/charles 3.04 2.7±0.01ms ? ?/sec 1.00 893.4±1.08µs ? ?/sec smol-songs.csv: desc + default/charles mingus 6.77 10.3±0.01ms ? ?/sec 1.00 1520.8±1.90µs ? ?/sec smol-songs.csv: desc + default/david 1.39 5.7±0.00ms ? ?/sec 1.00 4.1±0.00ms ? ?/sec smol-songs.csv: desc + default/david bowie 2.34 15.8±0.02ms ? ?/sec 1.00 6.7±0.01ms ? ?/sec smol-songs.csv: desc + default/john 1.00 2.5±0.00ms ? ?/sec 1.02 2.6±0.01ms ? ?/sec smol-songs.csv: desc + default/marcus miller 5.06 14.5±0.02ms ? ?/sec 1.00 2.9±0.01ms ? ?/sec smol-songs.csv: desc + default/michael jackson 2.64 14.1±0.05ms ? ?/sec 1.00 5.4±0.00ms ? ?/sec smol-songs.csv: desc + default/tamo 1.00 567.0±0.65µs ? ?/sec 1.00 565.7±0.97µs ? ?/sec smol-songs.csv: desc + default/thelonious monk 3.55 11.6±0.02ms ? ?/sec 1.00 3.3±0.00ms ? ?/sec smol-songs.csv: desc/Notstandskomitee 2.58 5.6±0.02ms ? ?/sec 1.00 2.2±0.02ms ? ?/sec smol-songs.csv: desc/charles 6.04 2.1±0.00ms ? ?/sec 1.00 348.1±0.57µs ? ?/sec smol-songs.csv: desc/charles mingus 14.51 9.4±0.01ms ? ?/sec 1.00 646.7±0.99µs ? ?/sec smol-songs.csv: desc/david 3.86 2.4±0.00ms ? ?/sec 1.00 620.7±2.46µs ? ?/sec smol-songs.csv: desc/david bowie 10.10 9.8±0.01ms ? ?/sec 1.00 973.9±3.31µs ? ?/sec smol-songs.csv: desc/john 1.00 545.5±0.78µs ? ?/sec 1.00 547.2±0.48µs ? ?/sec smol-songs.csv: desc/marcus miller 11.39 10.3±0.01ms ? ?/sec 1.00 903.7±0.95µs ? ?/sec smol-songs.csv: desc/michael jackson 10.51 9.7±0.01ms ? ?/sec 1.00 924.7±2.02µs ? ?/sec smol-songs.csv: desc/tamo 1.01 43.2±0.33µs ? ?/sec 1.00 42.6±0.35µs ? ?/sec smol-songs.csv: desc/thelonious monk 4.19 10.8±0.03ms ? ?/sec 1.00 2.6±0.00ms ? ?/sec smol-songs.csv: prefix search/a 1.00 1008.7±1.00µs ? ?/sec 1.00 1005.5±0.91µs ? ?/sec smol-songs.csv: prefix search/b 1.00 885.0±0.70µs ? ?/sec 1.01 890.6±1.11µs ? ?/sec smol-songs.csv: prefix search/i 1.00 1051.8±1.25µs ? ?/sec 1.00 1056.6±4.12µs ? ?/sec smol-songs.csv: prefix search/s 1.00 724.7±1.77µs ? ?/sec 1.00 721.6±0.59µs ? ?/sec smol-songs.csv: prefix search/x 1.01 212.4±0.21µs ? ?/sec 1.00 210.9±0.38µs ? ?/sec smol-songs.csv: proximity/7000 Danses Un Jour Dans Notre Vie 18.55 48.5±0.09ms ? ?/sec 1.00 2.6±0.03ms ? ?/sec smol-songs.csv: proximity/The Disneyland Sing-Along Chorus 8.41 56.7±0.45ms ? ?/sec 1.00 6.7±0.05ms ? ?/sec smol-songs.csv: proximity/Under Great Northern Lights 15.74 38.9±0.14ms ? ?/sec 1.00 2.5±0.00ms ? ?/sec smol-songs.csv: proximity/black saint sinner lady 11.82 40.1±0.13ms ? ?/sec 1.00 3.4±0.02ms ? ?/sec smol-songs.csv: proximity/les dangeureuses 1960 6.90 26.1±0.13ms ? ?/sec 1.00 3.8±0.04ms ? ?/sec smol-songs.csv: typo/Arethla Franklin 14.93 5.8±0.01ms ? ?/sec 1.00 390.1±1.89µs ? ?/sec smol-songs.csv: typo/Disnaylande 3.18 7.3±0.01ms ? ?/sec 1.00 2.3±0.00ms ? ?/sec smol-songs.csv: typo/dire straights 5.55 15.2±0.02ms ? ?/sec 1.00 2.7±0.00ms ? ?/sec smol-songs.csv: typo/fear of the duck 28.03 20.0±0.03ms ? ?/sec 1.00 713.3±1.54µs ? ?/sec smol-songs.csv: typo/indochie 19.25 1851.4±2.38µs ? ?/sec 1.00 96.2±0.13µs ? ?/sec smol-songs.csv: typo/indochien 14.66 1887.7±3.18µs ? ?/sec 1.00 128.8±0.18µs ? ?/sec smol-songs.csv: typo/klub des loopers 37.73 18.0±0.02ms ? ?/sec 1.00 476.7±0.73µs ? ?/sec smol-songs.csv: typo/michel depech 10.17 5.8±0.01ms ? ?/sec 1.00 565.8±1.16µs ? ?/sec smol-songs.csv: typo/mongus 15.33 1897.4±3.44µs ? ?/sec 1.00 123.8±0.13µs ? ?/sec smol-songs.csv: typo/stromal 14.63 1859.3±2.40µs ? ?/sec 1.00 127.1±0.29µs ? ?/sec smol-songs.csv: typo/the white striper 10.83 9.4±0.01ms ? ?/sec 1.00 866.0±0.98µs ? ?/sec smol-songs.csv: typo/thelonius monk 14.40 3.8±0.00ms ? ?/sec 1.00 261.5±1.30µs ? ?/sec smol-songs.csv: words/7000 Danses / Le Baiser / je me trompe de mots 5.54 70.8±0.09ms ? ?/sec 1.00 12.8±0.03ms ? ?/sec smol-songs.csv: words/Bring Your Daughter To The Slaughter but now this is not part of the title 3.48 119.8±0.14ms ? ?/sec 1.00 34.4±0.04ms ? ?/sec smol-songs.csv: words/The Disneyland Children's Sing-Alone song 8.98 71.9±0.12ms ? ?/sec 1.00 8.0±0.01ms ? ?/sec smol-songs.csv: words/les liaisons dangeureuses 1793 11.88 37.4±0.07ms ? ?/sec 1.00 3.1±0.01ms ? ?/sec smol-songs.csv: words/seven nation mummy 22.86 23.4±0.04ms ? ?/sec 1.00 1024.8±1.57µs ? ?/sec smol-songs.csv: words/the black saint and the sinner lady and the good doggo 2.76 124.4±0.15ms ? ?/sec 1.00 45.1±0.09ms ? ?/sec smol-songs.csv: words/whathavenotnsuchforth and a good amount of words to pop to match the first one 2.52 107.0±0.23ms ? ?/sec 1.00 42.4±0.66ms ? ?/sec group main-wiki typo-wiki ----- --------- --------- smol-wiki-articles.csv: basic placeholder/ 1.02 13.7±0.02µs ? ?/sec 1.00 13.4±0.03µs ? ?/sec smol-wiki-articles.csv: basic with quote/"film" 1.02 409.8±0.67µs ? ?/sec 1.00 402.6±0.48µs ? ?/sec smol-wiki-articles.csv: basic with quote/"france" 1.00 325.9±0.91µs ? ?/sec 1.00 326.4±0.49µs ? ?/sec smol-wiki-articles.csv: basic with quote/"japan" 1.00 218.4±0.26µs ? ?/sec 1.01 220.5±0.20µs ? ?/sec smol-wiki-articles.csv: basic with quote/"machine" 1.00 143.0±0.12µs ? ?/sec 1.04 148.8±0.21µs ? ?/sec smol-wiki-articles.csv: basic with quote/"miles" "davis" 1.00 11.7±0.06ms ? ?/sec 1.00 11.8±0.01ms ? ?/sec smol-wiki-articles.csv: basic with quote/"mingus" 1.00 4.4±0.03ms ? ?/sec 1.00 4.4±0.00ms ? ?/sec smol-wiki-articles.csv: basic with quote/"rock" "and" "roll" 1.00 43.5±0.08ms ? ?/sec 1.01 43.8±0.06ms ? ?/sec smol-wiki-articles.csv: basic with quote/"spain" 1.00 137.3±0.35µs ? ?/sec 1.05 144.4±0.23µs ? ?/sec smol-wiki-articles.csv: basic without quote/film 1.00 125.3±0.30µs ? ?/sec 1.06 133.1±0.37µs ? ?/sec smol-wiki-articles.csv: basic without quote/france 1.21 1782.6±1.65µs ? ?/sec 1.00 1477.0±1.39µs ? ?/sec smol-wiki-articles.csv: basic without quote/japan 1.28 1363.9±0.80µs ? ?/sec 1.00 1064.3±1.79µs ? ?/sec smol-wiki-articles.csv: basic without quote/machine 1.73 760.3±0.81µs ? ?/sec 1.00 439.6±0.75µs ? ?/sec smol-wiki-articles.csv: basic without quote/miles davis 1.03 17.0±0.03ms ? ?/sec 1.00 16.5±0.02ms ? ?/sec smol-wiki-articles.csv: basic without quote/mingus 1.07 5.3±0.01ms ? ?/sec 1.00 5.0±0.00ms ? ?/sec smol-wiki-articles.csv: basic without quote/rock and roll 1.01 63.9±0.18ms ? ?/sec 1.00 63.0±0.07ms ? ?/sec smol-wiki-articles.csv: basic without quote/spain 2.07 667.4±0.93µs ? ?/sec 1.00 322.8±0.29µs ? ?/sec smol-wiki-articles.csv: prefix search/c 1.00 343.1±0.47µs ? ?/sec 1.00 344.0±0.34µs ? ?/sec smol-wiki-articles.csv: prefix search/g 1.00 374.4±3.42µs ? ?/sec 1.00 374.1±0.44µs ? ?/sec smol-wiki-articles.csv: prefix search/j 1.00 359.9±0.31µs ? ?/sec 1.00 361.2±0.79µs ? ?/sec smol-wiki-articles.csv: prefix search/q 1.01 102.0±0.12µs ? ?/sec 1.00 101.4±0.32µs ? ?/sec smol-wiki-articles.csv: prefix search/t 1.00 536.7±1.39µs ? ?/sec 1.00 534.3±0.84µs ? ?/sec smol-wiki-articles.csv: prefix search/x 1.00 400.9±1.00µs ? ?/sec 1.00 399.5±0.45µs ? ?/sec smol-wiki-articles.csv: proximity/april paris 3.86 14.4±0.01ms ? ?/sec 1.00 3.7±0.01ms ? ?/sec smol-wiki-articles.csv: proximity/diesel engine 12.98 10.4±0.01ms ? ?/sec 1.00 803.5±1.13µs ? ?/sec smol-wiki-articles.csv: proximity/herald sings 1.00 12.7±0.06ms ? ?/sec 5.29 67.1±0.09ms ? ?/sec smol-wiki-articles.csv: proximity/tea two 6.48 1452.1±2.78µs ? ?/sec 1.00 224.1±0.38µs ? ?/sec smol-wiki-articles.csv: typo/Disnaylande 3.89 8.5±0.01ms ? ?/sec 1.00 2.2±0.01ms ? ?/sec smol-wiki-articles.csv: typo/aritmetric 3.78 10.3±0.01ms ? ?/sec 1.00 2.7±0.00ms ? ?/sec smol-wiki-articles.csv: typo/linax 8.91 1426.7±0.97µs ? ?/sec 1.00 160.1±0.18µs ? ?/sec smol-wiki-articles.csv: typo/migrosoft 7.48 1417.3±5.84µs ? ?/sec 1.00 189.5±0.88µs ? ?/sec smol-wiki-articles.csv: typo/nympalidea 3.96 7.2±0.01ms ? ?/sec 1.00 1810.1±2.03µs ? ?/sec smol-wiki-articles.csv: typo/phytogropher 3.71 7.2±0.01ms ? ?/sec 1.00 1934.3±6.51µs ? ?/sec smol-wiki-articles.csv: typo/sisan 6.44 1497.2±1.38µs ? ?/sec 1.00 232.7±0.94µs ? ?/sec smol-wiki-articles.csv: typo/the fronce 6.92 2.9±0.00ms ? ?/sec 1.00 418.0±1.76µs ? ?/sec smol-wiki-articles.csv: words/Abraham machin 16.63 10.8±0.01ms ? ?/sec 1.00 649.7±1.08µs ? ?/sec smol-wiki-articles.csv: words/Idaho Bellevue pizza 27.15 25.6±0.03ms ? ?/sec 1.00 944.2±5.07µs ? ?/sec smol-wiki-articles.csv: words/Kameya Tokujirō mingus monk 26.87 40.7±0.05ms ? ?/sec 1.00 1515.3±2.73µs ? ?/sec smol-wiki-articles.csv: words/Ulrich Hensel meilisearch milli 11.99 48.8±0.10ms ? ?/sec 1.00 4.1±0.02ms ? ?/sec smol-wiki-articles.csv: words/the black saint and the sinner lady and the good doggo 4.90 110.0±0.15ms ? ?/sec 1.00 22.4±0.03ms ? ?/sec ``` Co-authored-by: mpostma <postma.marin@protonmail.com> Co-authored-by: ad hoc <postma.marin@protonmail.com>	2022-03-15 16:43:36 +00:00
ad hoc	3f24555c3d	custom fst automatons	2022-03-15 17:38:35 +01:00
ad hoc	628c835a22	fix tests	2022-03-15 17:38:34 +01:00
bors[bot]	8efac33b53	Merge #467 467: optimize prefix database r=Kerollmops a=MarinPostma This pr introduces two optimizations that greatly improve the speed of computing prefix databases. - The time that it takes to create the prefix FST has been divided by 5 by inverting the way we iterated over the words FST. - We unconditionally and needlessly checked for documents to remove in `word_prefix_pair`, which caused an iteration over the whole database. Co-authored-by: ad hoc <postma.marin@protonmail.com>	2022-03-15 16:14:35 +00:00
ad hoc	d127c57f2d	review edits	2022-03-15 17:12:48 +01:00
ad hoc	d633ac5b9d	optimize word prefix pair	2022-03-15 16:37:22 +01:00
ad hoc	d68fe2b3c7	optimize word prefix fst	2022-03-15 16:36:48 +01:00
Kerollmops	08a06b49f0	Bump version to 0.23.1	2022-03-15 15:50:28 +01:00
bors[bot]	d87e8b63a9	Merge #465 465: Update dependencies r=ManyTheFish a=Kerollmops This PR upgrade and updates this crate's dependencies but first, it removes three dependencies that we don't use anymore. I used [cargo udeps](https://github.com/est31/cargo-udeps) to upgrade them ⬆️ Co-authored-by: Kerollmops <clement@meilisearch.com> Co-authored-by: Clément Renault <clement@meilisearch.com>	2022-03-15 13:49:17 +00:00
Clément Renault	0c5f4ed7de	Apply suggestions Co-authored-by: Many <many@meilisearch.com>	2022-03-15 14:18:29 +01:00
Kerollmops	21ec334dcc	Fix the compilation error of the dependency versions	2022-03-15 11:17:45 +01:00
Kerollmops	63682c2c9a	Upgrade the dependencies	2022-03-15 11:17:44 +01:00
Kerollmops	288a879411	Remove three useless dependencies	2022-03-15 11:17:44 +01:00
bors[bot]	712bf035a7	Merge #464 464: exporting heed to avoid having different versions of Heed in Meilisearch r=curquiza a=psvnlsaikumar # Pull Request ## What does this PR do? Fixes the issue in meilisearch https://github.com/meilisearch/meilisearch/issues/2210 ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: psvnl sai kumar <psvnlsaikumar@gmail.com>	2022-03-15 09:51:56 +00:00
psvnl sai kumar	5e08fac729	fixes for rustfmt pass	2022-03-14 19:22:41 +05:30
psvnl sai kumar	92e2e09434	exporting heed to avoid having different versions of Heed in Meilisearch	2022-03-14 01:01:58 +05:30
bors[bot]	290a29b5fb	Merge #457 457: Avoid iterating on big databases when useless r=Kerollmops a=Kerollmops This PR makes the prefix database updates to avoid iterating on big grenad files when it is unnecessary. We introduced this regression in #436 but it went unnoticed. --- According to the following benchmark results, we take more time when we index documents in one run than before #436. It looks like it is probably due to the fact that, now, instead of computing the prefixes database by iterating on the LMDB we directly iterate on the grenad file. Those could be slower to iterate on and could be the slowdown cause. I just pushed a commit that tests this branch with the new unreleased version of grenad where some work was done to speed up the iteration on grenad files. [The benchmarks for this last commit](https://github.com/meilisearch/milli/actions/runs/1927187408) are currently running. You can [see the diff](https://github.com/meilisearch/grenad/compare/v0.4.1...main) between the v0.4 and the unreleased v0.5 version of grenad. ```diff group indexing_benchmark-multi-batch-indexing-before-speed-up_45f52620 indexing_stop-iterating-on-big-grenad-files_ac8b85c4 ----- ---------------------------------------------------------------- ---------------------------------------------------- + indexing/Indexing songs in three batches with default settings 1.12 57.7±2.14s ? ?/sec 1.00 51.3±2.76s ? ?/sec - indexing/Indexing wiki 1.00 917.3±30.01s ? ?/sec 1.10 1008.4±38.27s ? ?/sec + indexing/Indexing wiki in three batches 1.10 1091.2±32.73s ? ?/sec 1.00 995.5±24.33s ? ?/sec ``` Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-03-09 16:46:34 +00:00
Kerollmops	1ae13c1374	Avoid iterating on big databases when useless	2022-03-09 15:43:54 +01:00
bors[bot]	a8d28e364d	Merge #461 461: Add a new error message when the `valid_fields` is empty r=curquiza a=brunoocasali I've created a test case to handle the new error formatting behavior, but I'm not sure if: - this is the right place to add the test? - this is the best way to test this behavior? And I'm not sure also regarding the `match` implementation, is this something required? Or maybe just an `if` statement is ok as well? I left the two messages literally without "reusing the prefix" in the implementation because I think this could help the "searchability" of the error in the future. # Pull Request ## What does this PR do? Fixes https://github.com/meilisearch/meilisearch/issues/2140 ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue? - [ ] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? Thank you so much for contributing to Meilisearch! Co-authored-by: Bruno Casali <brunoocasali@gmail.com>	2022-03-08 09:55:58 +00:00
bors[bot]	2ef5751795	Merge #463 463: Allow setting the primary-key in the cli r=irevoire a=irevoire Co-authored-by: Tamo <tamo@meilisearch.com>	2022-03-07 14:11:40 +00:00
Tamo	8bb45956d4	allow to set the primary key in the cli	2022-03-07 14:56:49 +01:00
bors[bot]	3cbadf92b6	Merge #462 462: cli improvements r=Kerollmops a=MarinPostma a few improvements: - use bufreader to load documents, so the loading of the document doesn't appear on flamegraphs - set default db path to current directory so the `-i` flag can be omitted. Co-authored-by: ad hoc <postma.marin@protonmail.com>	2022-03-07 09:39:01 +00:00
ad hoc	db3a1905de	default db path	2022-03-07 10:30:47 +01:00
ad hoc	6cf82ba993	bufread documents	2022-03-07 10:29:52 +01:00
Bruno Casali	66c6d5e1ef	Add a new error message when the `valid_fields` is empty > "Attribute `{}` is not sortable. This index doesn't have configured sortable attributes." > "Attribute `{}` is not sortable. Available sortable attributes are: `{}`." coexist in the error handling	2022-03-05 10:38:18 -03:00
bors[bot]	df518d8b0b	Merge #459 459: Update heed link in cargo toml r=Kerollmops a=curquiza Since grenad and heed have been moved to the meilisearch orga, this PR changes the link. This is a minor change since GitHub handles automatically the redirection. This PR is only for consisitency. Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-03-01 19:47:14 +00:00
Clémentine Urquizar	d9ed9de2b0	Update heed link in cargo toml	2022-03-01 19:45:29 +01:00
bors[bot]	51cf44d6fd	Merge #456 456: Remove useless grenad merging r=Kerollmops a=Kerollmops This PR must be merged after #454. This PR removes the part of code that was merging all of the grenad Readers merging that we don't need as the indexer should have merged them and, therefore, we should only have one final grenad Reader. We reduce the amount of CPU usage and memory pressure we were doing uselessly. `@ManyTheFish` are you sure I can skip merging the `word_docids` database? Here is the benchmark comparison with the previously merged PR #454: ``` group indexing_reintroduce-appending-sorted-values_c05e42a8 indexing_remove-useless-grenad-merging_d5b8b5a2 ----- ----------------------------------------------------- ----------------------------------------------- indexing/Indexing movies with default settings 1.06 16.6±1.04s ? ?/sec 1.00 15.7±0.93s ? ?/sec indexing/Indexing songs with default settings 1.16 60.1±7.07s ? ?/sec 1.00 51.7±5.98s ? ?/sec indexing/Indexing songs without faceted numbers 1.06 55.4±6.14s ? ?/sec 1.00 52.2±4.13s ? ?/sec ``` And the comparison with multi-batch indexing before #436, we can see that we gain time for benchmarks that index datasets in multiple batches but there is _so much_ variance that it's not clear. ``` group indexing_benchmark-multi-batch-indexing-before-speed-up_45f52620 indexing_remove-useless-grenad-merging_d5b8b5a2 ----- ---------------------------------------------------------------- ----------------------------------------------- indexing/Indexing geo_point 1.07 6.6±0.08s ? ?/sec 1.00 6.2±0.11s ? ?/sec indexing/Indexing songs in three batches with default settings 1.12 57.7±2.14s ? ?/sec 1.00 51.5±3.80s ? ?/sec indexing/Indexing songs with default settings 1.00 47.5±2.52s ? ?/sec 1.09 51.7±5.98s ? ?/sec indexing/Indexing songs without any facets 1.00 43.5±1.43s ? ?/sec 1.12 48.8±3.73s ? ?/sec indexing/Indexing songs without faceted numbers 1.00 47.1±2.23s ? ?/sec 1.11 52.2±4.13s ? ?/sec indexing/Indexing wiki 1.00 917.3±30.01s ? ?/sec 1.09 998.7±38.92s ? ?/sec indexing/Indexing wiki in three batches 1.09 1091.2±32.73s ? ?/sec 1.00 996.5±15.70s ? ?/sec ``` What do you think `@irevoire?` Should we change the benchmarks to make them do more runs? Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-03-01 16:48:08 +00:00
Kerollmops	d5b8b5a2f8	Replace the ugly unwraps by clean if let Somes	2022-02-28 16:31:33 +01:00
Kerollmops	8d26f3040c	Remove a useless grenad file merging	2022-02-28 16:31:33 +01:00
bors[bot]	21898ffc60	Merge #454 454: Reintroduce appending sorted entries when possible r=Kerollmops a=Kerollmops This PR modifies the `sorter_into_lmdb_database` function to append values into the database instead of get-put-merging them, it should improve the indexation speed for when the database is empty. ```txt group indexing_main_25123af3 indexing_reintroduce-appending-sorted-values_c05e42a8 ----- ---------------------- ----------------------------------------------------- indexing/Indexing movies with default settings 1.07 17.8±0.99s ? ?/sec 1.00 16.6±1.04s ? ?/sec indexing/Indexing songs with default settings 1.00 57.0±6.01s ? ?/sec 1.05 60.1±7.07s ? ?/sec indexing/Indexing songs without any facets 1.10 51.8±5.36s ? ?/sec 1.00 47.3±3.30s ? ?/sec ``` Co-authored-by: Clément Renault <clement@meilisearch.com>	2022-02-28 14:55:37 +00:00
Clément Renault	04b1bbf932	Reintroduce appending sorted entries when possible	2022-02-24 14:50:45 +01:00
bors[bot]	382be56d36	Merge #453 453: Benchmark multi batch indexing r=Kerollmops a=Kerollmops Hey `@irevoire,` could you please add the new benchmarks into influx? Co-authored-by: Kerollmops <clement@meilisearch.com> Co-authored-by: Clément Renault <clement@meilisearch.com>	2022-02-24 12:33:13 +00:00
Clément Renault	acfc96525c	Apply GitHub suggestions	2022-02-23 16:20:29 +01:00
Clément Renault	a820aa11e6	Add a new movies benchmark to test multi batch indexing	2022-02-23 16:20:29 +01:00
Kerollmops	8d2e3e4aba	Add a new wiki benchmark to test multi batch indexing	2022-02-23 16:20:29 +01:00
Kerollmops	ab5247dc64	Add a new songs benchmark to test multi batch indexing	2022-02-23 16:20:28 +01:00
bors[bot]	acd9535588	Merge #455 455: Raise the GitHub CI timeout limit to 72h r=irevoire a=Kerollmops Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-02-23 14:33:31 +00:00
Kerollmops	19bfb2649b	Raise the GitHub CI timeout limit to 72h	2022-02-23 15:27:51 +01:00
bors[bot]	25123af3b8	Merge #436 436: Speed up the word prefix databases computation time r=Kerollmops a=Kerollmops This PR depends on the fixes done in #431 and must be merged after it. In this PR we will bring the `WordPrefixPairProximityDocids`, `WordPrefixDocids` and, `WordPrefixPositionDocids` update structures to a new era, a better era, where computing the word prefix pair proximities costs much fewer CPU cycles, an era where this update structure can use the, previously computed, set of new word docids from the newly indexed batch of documents. --- The `WordPrefixPairProximityDocids` is an update structure, which means that it is an object that we feed with some parameters and which modifies the LMDB database of an index when asked for. This structure specifically computes the list of word prefix pair proximities, which correspond to a list of pairs of words associated with a proximity (the distance between both words) where the second word is not a word but a prefix e.g. `s`, `se`, `a`. This word prefix pair proximity is associated with the list of documents ids which contains the pair of words and prefix at the given proximity. The origin of the performances issue that this struct brings is related to the fact that it starts its job from the beginning, it clears the LMDB database before rewriting everything from scratch, using the other LMDB databases to achieve that. I hope you understand that this is absolutely not an optimized way of doing things. Co-authored-by: Clément Renault <clement@meilisearch.com> Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-02-16 15:41:14 +00:00
Clément Renault	ff8d7a810d	Change the behavior of the as_cloneable_grenad by taking a ref	2022-02-16 15:40:08 +01:00
Clément Renault	f367cc2e75	Finally bump grenad to v0.4.1	2022-02-16 15:28:48 +01:00
bors[bot]	f2984f66e6	Merge #452 452: bump milli r=curquiza a=irevoire Co-authored-by: Irevoire <tamo@meilisearch.com>	2022-02-16 13:49:14 +00:00
Irevoire	0defeb268c	bump milli	2022-02-16 13:27:41 +01:00
bors[bot]	030064da25	Merge #451 451: Update LICENSE with Meili SAS name r=Kerollmops a=curquiza Check with thomas, we must put the real name of the company Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com>	2022-02-15 16:18:47 +00:00
Clémentine Urquizar - curqui	84035a27f5	Update LICENSE	2022-02-15 15:52:50 +01:00
bors[bot]	0885fcf973	Merge #450 450: Get rid of chrono in favor of time r=Kerollmops a=irevoire We only use `chrono` as a wrapper around `time`, and since there has been an [open CVE on `chrono` for at least 3 months now](https://github.com/chronotope/chrono/pull/632) and the repo seems to be [struggling with maintenance](https://github.com/chronotope/chrono/pull/639), I think we should use `time` directly which is way more active and sufficient for our use case. EDIT: Actually the CVE status has been known for more than 6 months: https://github.com/chronotope/chrono/issues/602 Co-authored-by: Irevoire <tamo@meilisearch.com>	2022-02-15 10:54:46 +00:00
Irevoire	48542ac8fd	get rid of chrono in favor of time	2022-02-15 11:41:55 +01:00
bors[bot]	ea15ad6c34	Merge #447 447: Update version for the next release (v0.22.1) r=curquiza a=curquiza Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-02-07 17:44:09 +00:00
Clémentine Urquizar	d03b3ceb58	Update version for the next release (v0.22.1)	2022-02-07 18:39:29 +01:00
bors[bot]	5d58cb7449	Merge #442 442: fix phrase search r=curquiza a=MarinPostma Run the exact match search on 7 words windows instead of only two. This makes false positive very very unlikely, and impossible on phrase query that are less than seven words. Co-authored-by: ad hoc <postma.marin@protonmail.com>	2022-02-07 16:18:20 +00:00
bors[bot]	c5a996aa78	Merge #446 446: Update LICENSE r=Kerollmops a=curquiza Co-authored-by: Clémentine Urquizar - curqui <clementine@meilisearch.com>	2022-02-07 09:47:39 +00:00
Clémentine Urquizar - curqui	1279c38ac9	Update LICENSE	2022-02-05 18:29:11 +01:00
bors[bot]	267d14c28d	Merge #445 445: allow null values in csv r=Kerollmops a=MarinPostma This pr allows null values in csv: - if the field is of type string, then an empty field is considered null (`,,`), anything other is turned into a string (i.e `, ,` is a single whitespace string) - if the field is of type number, when the trimmed field is empty, we consider the value null (i.e `,,`, `, ,` are both null numbers) otherwise we try to parse the number. Co-authored-by: ad hoc <postma.marin@protonmail.com>	2022-02-03 15:11:32 +00:00
ad hoc	bd2262ceea	allow null values in csv	2022-02-03 16:03:01 +01:00
ad hoc	13de251047	rewrite word pair distance gathering	2022-02-03 15:57:20 +01:00
bors[bot]	fda4f229bb	Merge #417 417: Change chunk size to 4MiB to fit more the end user usage r=Kerollmops a=ManyTheFish Reverts meilisearch/milli#379 We made several indexing tests using different sizes of datasets (5 datasets from 9MiB to 100MiB) on several typologies of VMs (`XS: 1GiB RAM, 1 VCPU`, `S: 2GiB RAM, 2 VCPU`, `M: 4GiB RAM, 3 VCPU`, `L: 8GiB RAM, 4 VCPU`). The result of these tests shows that the `4MiB` chunk size seems to be the best size compared to other chunk sizes (`2Mib`, `4MiB`, `8Mib`, `16Mib`, `32Mib`, `64Mib`, `128Mib`). below is the average time per chunk size: ![Capture d’écran 2021-09-27 à 14 27 50](https://user-images.githubusercontent.com/6482087/134909368-ef0bc45e-68d5-49d1-aaf9-91113b7c410f.png) <details> <summary>Detailled data</summary> <br> ![Capture d’écran 2021-09-27 à 14 39 48](https://user-images.githubusercontent.com/6482087/134909952-a36b1457-bbbd-4a6c-bbe5-519e4b926b5a.png) </br> </details> Co-authored-by: Many <many@meilisearch.com>	2022-02-02 18:30:59 +00:00
bors[bot]	2468ebb76b	Merge #444 444: Fix the parsing of ndjson requests to index more than the first line r=Kerollmops a=Kerollmops This PR correctly uses the `BufRead` trait to read every line of the content instead of just the first one. This bug was only affecting the http-ui test crate. Co-authored-by: Kerollmops <clement@meilisearch.com>	2022-02-02 17:59:44 +00:00
Kerollmops	9142ba9dd4	Fix the parsing of ndjson requests to index more than the first line	2022-02-02 17:55:13 +01:00
Many	d59bcea749	Revert "Revert "Change chunk size to 4MiB to fit more the end user usage""	2022-02-02 17:01:13 +01:00
mpostma	7541ab99cd	review changes	2022-02-02 12:59:01 +01:00
mpostma	d0aabde502	optimize 2 typos case	2022-02-02 12:56:09 +01:00
mpostma	55e6cb9c7b	typos on first letter counts as 2	2022-02-02 12:56:09 +01:00
mpostma	642c01d0dc	set max typos on ngram to 1	2022-02-02 12:56:08 +01:00
ad hoc	d852dc0d2b	fix phrase search	2022-02-01 20:21:33 +01:00
Kerollmops	fb79c32430	Compute the new, common and, deleted prefix words fst once	2022-01-27 11:00:18 +01:00
Clément Renault	51d1e64b23	Remove, now useless, the WriteMethod enum	2022-01-27 10:08:35 +01:00
Clément Renault	e9c02173cf	Rework the WordsPrefixPositionDocids update to compute a subset of the database	2022-01-27 10:08:35 +01:00
Clément Renault	dbba5fd461	Create a function to simplify the word prefix pair proximity docids compute	2022-01-27 10:08:35 +01:00
Clément Renault	e760e02737	Fix the computation of the newly added and common prefix pair proximity words	2022-01-27 10:08:35 +01:00
Clément Renault	d59e559317	Fix the computation of the newly added and common prefix words	2022-01-27 10:08:34 +01:00
Clément Renault	2ec8542105	Rework the WordPrefixDocids update to compute a subset of the database	2022-01-27 10:08:34 +01:00
Clément Renault	28692f65be	Rework the WordPrefixDocids update to compute a subset of the database	2022-01-27 10:08:34 +01:00
Clément Renault	5404bc02dd	Move the fst_stream_into_hashset method in the helper methods	2022-01-27 10:06:00 +01:00
Clément Renault	c90fa95f93	Only compute the word prefix pairs on the created word pair proximities	2022-01-27 10:06:00 +01:00
Clément Renault	822f67e9ad	Bring the newly created word pair proximity docids	2022-01-27 10:06:00 +01:00
Clément Renault	d28f18658e	Retrieve the previous version of the words prefixes FST	2022-01-27 10:05:59 +01:00
bors[bot]	38d23546a5	Merge #431 431: Fix and improve word prefix pair proximity r=ManyTheFish a=Kerollmops This PR first fixes the algorithm we used to select and compute the word prefix pair proximity database. The previous version was skipping nearly all of the prefixes. The issue is that this fix made this method to take more time and we were trying to reduce the time spent in it. With `@ManyTheFish` we found out that we could skip some of the work we were doing by: - discarding the prefixes that were shorter than a specific threshold (default: 2). - discarding the word prefix pairs with proximity bigger than a specific threshold (default: 4). - remove the unused threshold that was specifying a minimum amount of word docids to merge. We will take more time to do some more optimization, like stop clearing and recomputing from scratch the database, we will compute the subsets of keys to create, keep and merge. This change is a little bit more complex than what this PR does. I keep this PR as a draft as I want to further test the real gain if it is enough or not if it is valid or not. I advise reviewers to review commit by commit to see the changes bit by bit, reviewing the whole PR can be hard. Co-authored-by: Clément Renault <clement@meilisearch.com>	2022-01-27 07:04:56 +00:00
bors[bot]	c63f945093	Merge #441 441: Changes related to the rebranding r=curquiza a=meili-bot _This PR is auto-generated._ - [X] Change the name `MeiliSearch` to `Meilisearch` in README. - [x] ⚠️ Ensure the bot did not update part you don’t want it to update, especially in the code examples in the Getting started. - [x] Please, ensure there is no other "MeiliSearch". For example, in the comments or in the tests name. - [x] Put the new logo on the README if needed -> still using the milli logo so far Co-authored-by: meili-bot <74670311+meili-bot@users.noreply.github.com> Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-01-26 17:07:37 +00:00
Clémentine Urquizar	0f213f2202	Replace MeiliSearch by Meilisearch	2022-01-26 17:49:55 +01:00
Clémentine Urquizar	de808a391a	Replace meilisearch by Meilisearch	2022-01-26 17:48:22 +01:00
meili-bot	0d282e3cc5	Update README.md	2022-01-26 16:33:16 +01:00
bors[bot]	d342c3c357	Merge #438 438: CLI improvements r=Kerollmops a=MarinPostma I've made the following changes to the cli: - `settings-update` become `settings`, with two subcommands: `update` and `show`. - `document-addition` becomes `documents` with a subcommands: `add` (I'll add a feature to list documents later) - `search` now has an interactive mode `-i` - search return the number of documents and the time it took to perform the search. Co-authored-by: mpostma <postma.marin@protonmail.com>	2022-01-26 15:18:20 +00:00
Clément Renault	f9b214f34e	Apply suggestions from code review Co-authored-by: Many <legendre.maxime.isn@gmail.com>	2022-01-26 11:28:11 +01:00
bors[bot]	e1cc025cbd	Merge #440 440: fix(fuzzer): fix the fuzzer after #430 r=Kerollmops a=irevoire Co-authored-by: Tamo <tamo@meilisearch.com>	2022-01-25 16:33:57 +00:00
Clément Renault	f04cd19886	Introduce a max prefix length parameter to the word prefix pair proximity update	2022-01-25 17:04:23 +01:00
Clément Renault	1514dfa1b7	Introduce a max proximity parameter to the word prefix pair proximity update	2022-01-25 17:04:23 +01:00
Clément Renault	23ea3ad738	Remove the useless threshold when computing the word prefix pair proximity	2022-01-25 17:04:23 +01:00
Clément Renault	e3c34684c6	Fix a bug where we were skipping most of the prefix pairs	2022-01-25 17:04:23 +01:00
mpostma	b5f01b52c7	cli improvements	2022-01-25 14:08:30 +01:00
Tamo	fb51d511be	fix(fuzzer): fix the fuzzer after #430	2022-01-25 12:08:47 +01:00
bors[bot]	9f2ff71581	Merge #434 434: bump milli to v0.22.0 r=curquiza a=irevoire This is breaking because of this PR: `98a365aaae` Should we do a special branch to only release the [patch](https://github.com/meilisearch/milli/pull/433) for https://github.com/meilisearch/MeiliSearch/issues/2082 (which is non-breaking)? Co-authored-by: Tamo <tamo@meilisearch.com>	2022-01-24 17:31:20 +00:00
bors[bot]	fd177b63f8	Merge #423 423: Remove an unused file r=irevoire a=irevoire This empty file is not included anywhere Co-authored-by: Tamo <tamo@meilisearch.com>	2022-01-19 14:18:05 +00:00
bors[bot]	8433516d85	Merge #430 430: Document batch support r=Kerollmops a=MarinPostma This pr adds support for document batches in milli. It changes the API of the `IndexDocuments` builder by adding a `add_documents` method. The API of the updates is changed a little, with the `UpdateBuilder` being renamed to `IndexerConfig` and being passed to the update builders. This makes it easier to pass around structs that need to access the indexer config, rather that extracting the fields each time. This change impacts many function signatures and simplify them. The change in not thorough, and may require another PR to propagate to the whole codebase. I restricted to the necessary for this PR. Co-authored-by: Marin Postma <postma.marin@protonmail.com>	2022-01-19 13:32:59 +00:00
Marin Postma	0c84a40298	document batch support reusable transform rework update api add indexer config fix tests review changes Co-authored-by: Clément Renault <clement@meilisearch.com> fmt	2022-01-19 12:40:20 +01:00
bors[bot]	74962b2fd9	Merge #435 435: Ensure we get no documents and no error when filtering on an empty db r=Kerollmops a=irevoire Co-authored-by: Tamo <tamo@meilisearch.com>	2022-01-18 10:46:26 +00:00
Tamo	01968d7ca7	ensure we get no documents and no error when filtering on an empty db	2022-01-18 11:40:30 +01:00
Tamo	367f403693	bump milli	2022-01-17 16:41:34 +01:00
bors[bot]	8f4499090b	Merge #433 433: fix(filter): Fix two bugs. r=Kerollmops a=irevoire - Stop lowercasing the field when looking in the field id map - When a field id does not exist it means there is currently zero documents containing this field thus we return an empty RoaringBitmap instead of throwing an internal error Will fix https://github.com/meilisearch/MeiliSearch/issues/2082 once meilisearch is released Co-authored-by: Tamo <tamo@meilisearch.com>	2022-01-17 14:06:53 +00:00
bors[bot]	4c516c00da	Merge #426 426: Fix search highlight for non-unicode chars r=ManyTheFish a=Samyak2 # Pull Request ## What does this PR do? Fixes https://github.com/meilisearch/MeiliSearch/issues/1480 <!-- Please link the issue you're trying to fix with this PR, if none then please create an issue first. --> ## PR checklist Please check if your PR fulfills the following requirements: - [x] Does this PR fix an existing issue? - [x] Have you read the contributing guidelines? - [x] Have you made sure that the title is accurate and descriptive of the changes? ## Changes The `matching_bytes` function takes a `&Token` now and: - gets the number of bytes to highlight (unchanged). - uses `Token.num_graphemes_from_bytes` to get the number of grapheme clusters to highlight. In essence, the `matching_bytes` function now returns the number of matching grapheme clusters instead of bytes. Added proper highlighting in the HTTP UI: - requires dependency on `unicode-segmentation` to extract grapheme clusters from tokens - `<mark>` tag is put around only the matched part - before this change, the entire word was highlighted even if only a part of it matched ## Questions Since `matching_bytes` does not return number of bytes but grapheme clusters, should it be renamed to something like `matching_chars` or `matching_graphemes`? Will this break the API? Thank you very much `@ManyTheFish` for helping 😄 Co-authored-by: Samyak S Sarnayak <samyak201@gmail.com>	2022-01-17 13:39:00 +00:00
Tamo	d1ac40ea14	fix(filter): Fix two bugs. - Stop lowercasing the field when looking in the field id map - When a field id does not exist it means there is currently zero documents containing this field thus we returns an empty RoaringBitmap instead of throwing an internal error	2022-01-17 13:51:46 +01:00
bors[bot]	15bbde1022	Merge #432 432: Fuzzer r=Kerollmops a=irevoire Provide a first way of fuzzing the indexing part of milli. It depends on [cargo-fuzz](https://rust-fuzz.github.io/book/cargo-fuzz.html) Co-authored-by: Tamo <tamo@meilisearch.com>	2022-01-17 12:50:26 +00:00
Samyak S Sarnayak	c0313f3026	Use chars for highlight instead of graphemes Tokenizer v0.2.7 uses chars instead of graphemes for matching bytes. `unicode-segmentation` dependency isn't needed anymore. Also, oxidised the highlight code :) Co-authored-by: many <maxime@meilisearch.com>	2022-01-17 13:15:31 +05:30
Samyak S Sarnayak	2d7607734e	Run cargo fmt on matching_words.rs	2022-01-17 13:04:33 +05:30
Samyak S Sarnayak	5ab505be33	Fix highlight by replacing num_graphemes_from_bytes num_graphemes_from_bytes has been renamed in the tokenizer to num_chars_from_bytes. Highlight now works correctly!	2022-01-17 13:02:55 +05:30
Samyak S Sarnayak	c10f58b7bd	Update tokenizer to v0.2.7	2022-01-17 13:02:00 +05:30
Samyak S Sarnayak	e752bd06f7	Fix matching_words tests to compile successfully The tests still fail due to a bug in https://github.com/meilisearch/tokenizer/pull/59	2022-01-17 11:37:45 +05:30
Samyak S Sarnayak	30247d70cd	Fix search highlight for non-unicode chars The `matching_bytes` function takes a `&Token` now and: - gets the number of bytes to highlight (unchanged). - uses `Token.num_graphemes_from_bytes` to get the number of grapheme clusters to highlight. In essence, the `matching_bytes` function returns the number of matching grapheme clusters instead of bytes. Should this function be renamed then? Added proper highlighting in the HTTP UI: - requires dependency on `unicode-segmentation` to extract grapheme clusters from tokens - `<mark>` tag is put around only the matched part - before this change, the entire word was highlighted even if only a part of it matched	2022-01-17 11:37:44 +05:30
Tamo	0605c0ac68	apply review comments	2022-01-13 18:51:08 +01:00
Tamo	b22c80106f	add some settings to the fuzzed milli and use the published version of arbitrary json	2022-01-13 15:35:24 +01:00
Tamo	c94952e25d	update the readme + dependencies	2022-01-12 18:30:11 +01:00
Tamo	e1053989c0	add a fuzzer on milli	2022-01-12 17:57:54 +01:00
bors[bot]	559e019de1	Merge #424 424: Store the geopoint in three dimensions r=Kerollmops a=irevoire Related to this issue: https://github.com/meilisearch/MeiliSearch/issues/1872 Fix the whole computation of distance for any “geo” operations (sort or filter). Now when you sort points they are returned to you in the right order. And when you filter on a specific radius you only get points included in the radius. This PR changes the way we store the geo points in the RTree. Instead of considering the latitude and longitude as orthogonal coordinates, we convert them to real orthogonal coordinates projected on a sphere with a radius of 1. This is the conversion formulae. ![image](https://user-images.githubusercontent.com/7032172/145990456-eefe840a-384f-4486-848b-81d0036814ec.png) Which, in rust, translate to this function: ```rust pub fn lat_lng_to_xyz(coord: &[f64; 2]) -> [f64; 3] { let [lat, lng] = coord.map(\|f\| f.to_radians()); let x = lat.cos() * lng.cos(); let y = lat.cos() * lng.sin(); let z = lat.sin(); [x, y, z] } ``` Storing the points on a sphere is easier / faster to compute than storing the point on an approximation of the real earth shape. But when we need to compute the distance between two points we still need to use the haversine distance which works with latitude and longitude. So, to do the fewest search-time computation possible I'm now associating every point with its `DocId` and its lat/lng. Co-authored-by: Tamo <tamo@meilisearch.com>	2022-01-10 15:23:43 +00:00
bors[bot]	660eac50b2	Merge #427 427: Handle escaped characters in filters r=Kerollmops a=irevoire Co-authored-by: Tamo <tamo@meilisearch.com>	2022-01-10 15:01:23 +00:00
Tamo	92804f6f45	apply clippy suggestions	2022-01-10 15:59:04 +01:00
Tamo	0fcde35a20	Update filter-parser/src/value.rs Co-authored-by: Clément Renault <clement@meilisearch.com>	2022-01-10 15:53:44 +01:00
Tamo	3c7ea1d298	Apply code suggestions Co-authored-by: Clément Renault <clement@meilisearch.com>	2022-01-10 15:19:21 +01:00
bors[bot]	74594be234	Merge #429 429: Benchmark CIs: not use a default label to call the GH runner r=irevoire a=curquiza Since we now have multiple self-hosted github runners, we need to differentiate them calling them in the CI. The `self-hosted` label is the default one, so we need to use the unique and appropriate one for the benchmark machine <img width="925" alt="Capture d’écran 2022-01-04 à 15 42 18" src="https://user-images.githubusercontent.com/20380692/148079840-49cd7878-5912-46ff-8ab8-bf646777f782.png"> Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2022-01-04 15:41:08 +00:00
Clémentine Urquizar	3d99686f7a	Change self-hosted label by benchmarks	2022-01-04 16:01:01 +01:00
bors[bot]	c039562723	Merge #428 428: Reintroduce the gitignore for the fuzzer r=Kerollmops a=irevoire Reintroduce the gitignore in the fuzz directory Co-authored-by: Tamo <tamo@meilisearch.com>	2022-01-04 12:09:06 +00:00
Tamo	9bdcd42b9b	reintroduce the gitignore for the fuzzer	2022-01-04 13:07:32 +01:00
bors[bot]	4cae691b86	Merge #425 425: Push the result of the benchmarks to influxdb r=irevoire a=irevoire Now execute a benchmark for every PR merged into main and then upload the results to influxdb. Co-authored-by: Tamo <tamo@meilisearch.com>	2022-01-04 11:04:16 +00:00
Tamo	6a1216bd51	Integrate telegraf into our CI	2022-01-04 11:59:05 +01:00
Tamo	02a21fd309	Handle the escapes of quote in the filters	2022-01-04 04:04:10 +01:00
Tamo	98a365aaae	store the geopoint in three dimensions	2021-12-14 12:21:24 +01:00
Tamo	d671d6f0f1	remove an unused file	2021-12-13 19:27:34 +01:00
bors[bot]	11a056d116	Merge #422 422: Prefer returning `None` instead of using an `FilterCondition::Empty` state r=Kerollmops a=Kerollmops This PR is related to the issue comment https://github.com/meilisearch/MeiliSearch/issues/1338#issuecomment-989322889 which exhibits the fact that when a filter is known to be empty no results are returned which is wrong, the filter should not apply as no restriction is done on the documents set. The filter system on the milli side has introduced an Empty state which was used in this kind of situation but I found out that it is not needed and that when we parse a filter and that it is empty we can simply return `None` as the `Filter::from_array` constructor does. So I removed it and added tests! On the MeiliSearch side, we just need to match on a `None` and completely ignore the filter in such a case. Co-authored-by: Clément Renault <clement@meilisearch.com>	2021-12-09 15:03:04 +00:00
Clément Renault	94011bb9a8	Fix the benchmarks to work with optional filters	2021-12-09 12:14:16 +01:00
Clément Renault	1c6c89f345	Fix the binaries that use the new optional filters	2021-12-09 11:57:53 +01:00
Clément Renault	25faef67d0	Remove the database setup in the filter_depth test	2021-12-09 11:57:53 +01:00
Clément Renault	65519bc04b	Test that empty filters return a None	2021-12-09 11:57:53 +01:00
Clément Renault	ef59762d8e	Prefer returning None instead of the Empty Filter state	2021-12-09 11:57:52 +01:00
bors[bot]	80dcfd5c3e	Merge #421 421: Introduce the depth method on FilterCondition r=Kerollmops a=Kerollmops This PR introduces the depth method on the FilterCondition type to be able to react to it. It is meant to be used to reject filters that go too deep and can make the engine stack overflow. Co-authored-by: Clément Renault <clement@meilisearch.com>	2021-12-09 10:28:52 +00:00
Clément Renault	ee856a7a46	Limit the max filter depth to 2000	2021-12-07 17:36:45 +01:00
Clément Renault	32bd9f091f	Detect the filters that are too deep and return an error	2021-12-07 17:20:11 +01:00
Clément Renault	90f49eab6d	Check the filter max depth limit and reject the invalid ones	2021-12-07 16:32:48 +01:00
Clément Renault	49c2db9485	Change the depth function to return the token depth	2021-12-07 16:06:10 +01:00
Clément Renault	57502fcf6a	Introduce the depth method on FilterCondition	2021-12-06 17:35:20 +01:00
bors[bot]	c83b77304a	Merge #420 420: Update milli 0.21.0 r=ManyTheFish a=ManyTheFish Update all modules to 0.21.0 Co-authored-by: many <maxime@meilisearch.com>	2021-11-30 17:22:12 +00:00
many	1b3923b5ce	Update all packages to 0.21.0	2021-11-29 12:17:59 +01:00
bors[bot]	26629a3f9e	Merge #419 419: fix word pair proximity indexing r=ManyTheFish a=ManyTheFish # Pull Request Sort positions before iterating over them during word pair proximity extraction. fixes [Meilisearch#1913](https://github.com/meilisearch/MeiliSearch/issues/1913) Co-authored-by: many <maxime@meilisearch.com>	2021-11-23 10:21:05 +00:00
many	8970246bc4	Sort positions before iterating over them during word pair proximity extraction	2021-11-22 18:16:54 +01:00
bors[bot]	cc32519a2d	Merge #418 418: change visibility of DocumentDeletionResult r=Kerollmops a=MarinPostma Change the visibility of `DocumentDeletionResult`, so its fields can be accesses from outside milli. Co-authored-by: Marin Postma <postma.marin@protonmail.com>	2021-11-22 14:45:55 +00:00
Marin Postma	6e977dd8e8	change visibility of DocumentDeletionResult	2021-11-22 15:44:44 +01:00
bors[bot]	68f1db123a	Merge #416 416: Update tokenizer v0.2.6 r=Kerollmops a=ManyTheFish Co-authored-by: many <maxime@meilisearch.com>	2021-11-18 16:01:11 +00:00
many	35f9499638	Export tokenizer from milli	2021-11-18 16:57:12 +01:00
many	64ef5869d7	Update tokenizer v0.2.6	2021-11-18 16:56:05 +01:00
bors[bot]	2c14efa8a2	Merge #409 409: remove update_id in UpdateBuilder r=ManyTheFish a=MarinPostma Removing the `update_id` from `UpdateBuidler`, since it serves no purpose. I had introduced it when working in HA some time ago, but I think there are better ways to do it now, so it can be removed an stop being in our way. Co-authored-by: Marin Postma <postma.marin@protonmail.com>	2021-11-16 14:59:09 +00:00
Marin Postma	6eb47ab792	remove update_id in UpdateBuilder	2021-11-16 13:07:04 +01:00
bors[bot]	21b78f3926	Merge #414 414: improve update result types r=ManyTheFish a=MarinPostma Inprove the returned meta when performing document additions and deletions: - On document addition return the number of indexed documents and the total number of documents in the index after the indexion - On document deletion return the number of deleted documents, and the remaining number of documents in the index after the deletion is performed I also fixed a potential bug when performing a document deletion and the primary key couldn't be found: before we assumed that the db was empty and returned that no documents were deleted, but since we checked before that the db wasn't empty, entering this branch is actually a bug, and now returns a 'MissingPrimaryKey' error. Co-authored-by: Marin Postma <postma.marin@protonmail.com>	2021-11-15 09:06:10 +00:00
Marin Postma	09b4281cff	improve document addition returned metaimprove document addition returned metaimprove document addition returned metaimprove document addition returned metaimprove document addition returned metaimprove document addition returned metaimprove document addition returned metaimprove document addition returned meta	2021-11-10 14:08:36 +01:00
Marin Postma	721fc294be	improve document deletion returned meta returns both the remaining number of documents and the number of deleted documents.	2021-11-10 14:08:18 +01:00
bors[bot]	8dff08d772	Merge #400 400: Rewrite the filter parser and add a lot of tests r=irevoire a=irevoire This PR is a complete rewrite of #358, which was reverted in #403. You can already try this PR in Meilisearch here https://github.com/meilisearch/MeiliSearch/pull/1880. Since writing a parser is quite complicated, I moved all the logic to another workspace called `filter_parser`. In this workspace, we don't know anything about milli, the filterable fields / field ID or anything. As you can see in its `cargo.toml`, it has only three dependencies entirely focused on the parsing part: ``` nom = "7.0.0" nom_locate = "4.0.0" ``` But introducing this new workspace made some changes necessary on the “AST”. Now the parser only returns `Tokens` (a simple `&str` with a bit of context). Everything is interpreted when we execute the filter later in milli. This crate provides a new error type for all filter related errors. --------- ## Errors Currently, we have multiple kinds of errors. Sometimes we are generating errors looking like that: (for `name = truc`) ``` Attribute `name` is not filterable. Available filterable attributes are: ``. ``` While sometimes pest was generating errors looking like that: ``` Invalid syntax for the filter parameter: ` --> 1:7 \| 1 \| name = \| ^--- \| = expected word`. ``` Which most people were seeing like that: (for `name =`) ``` Invalid syntax for the filter parameter: ` --> 1:7\n \|\n1 \| name =\n \| ^---\n \|\n = expected word`. ``` ----------- With this PR, the error format is unified between all errors. All errors follow this more straightforward format: ``` The error message. [from char]:[to char] filter ``` This should be way easier to read when embedded in the JSON for a human. And it should also allow us to parse the errors easily and provide highlighting or something with a frontend playground. Here is an example of the two previous errors with the new format: For `name = truc`: ``` Attribute `name` is not filterable. Available filterable attributes are: ``. 1:4 name = truc ``` Or in one line: ``` Attribute `name` is not filterable. Available filterable attributes are: ``.\n1:4 name = truc ``` And for `name =`: ``` Was expecting a value but instead got nothing. 7:7 name = ``` Or in one line: ``` Was expecting a value but instead got nothing.\n7:7 name = ``` Also, since we now have control over the parser, we can generate more explicit error messages so a lot of new errors have been created. I tried to be as helpful as possible for the user; here is a little overview of the new error message you can get when misusing a filter: ``` Expression `"truc` is missing the following closing delimiter: `"`. 8:13 name = "truc ``` The `_geoRadius` filter is an operation and can't be used as a value. 8:30 name = _geoRadius(12, 13, 14) ``` etc ## Tests A lot of tests have been written in the `filter_parser` crate. I think there is a unit test for every part of the syntax. But since we can never be sure we covered all the cases, I also fuzzed the new parser A LOT (for ±8 hours on 20 threads). And the code to fuzz the parser is included in the workspace, so if one day we need to change something to the syntax, we'll be able to re-use it by simply running: ``` cargo fuzz run --release parse ``` ## Milli I renamed the type and module `filter_condition.rs` / `FilterCondition` to `filter.rs` / `Filter`. Co-authored-by: Tamo <tamo@meilisearch.com>	2021-11-09 16:09:34 +00:00
Irevoire	7c3017734a	re-ignore the ! symbol when generating a good error message	2021-11-09 17:08:04 +01:00
Tamo	bff48681d2	Re-order the operator Co-authored-by: Clément Renault <clement@meilisearch.com>	2021-11-09 17:05:36 +01:00
Irevoire	519d6b2bf3	remove the `!` syntax for the not	2021-11-09 16:47:54 +01:00
Irevoire	73df873f44	fix typos	2021-11-09 16:41:10 +01:00
Irevoire	99197387af	fix the test with the new escaped format	2021-11-09 16:41:10 +01:00
Tamo	f28600031d	Rename the filter_parser crate into filter-parser Co-authored-by: Clément Renault <clement@meilisearch.com>	2021-11-09 16:41:10 +01:00
Irevoire	0ea0146e04	implement deref &str on the tokens	2021-11-09 11:34:10 +01:00
Irevoire	a211a9cdcd	update the error format so it can be easily parsed by someone else	2021-11-09 11:19:30 +01:00
Irevoire	9b24f83456	in case of error return a range of chars position instead of one line and column	2021-11-09 10:27:29 +01:00
Tamo	2c6d08c519	Simplify the tokens to only wrap one span and no inner value Co-authored-by: marin <postma.marin@protonmail.com>	2021-11-09 10:12:20 +01:00
Irevoire	18eb4b9c51	fix spaces in the bnf	2021-11-09 01:04:50 +01:00
Tamo	cf98bf37d0	Simplify some closure Co-authored-by: marin <postma.marin@protonmail.com>	2021-11-09 01:03:02 +01:00
Tamo	bc9daf9041	update the bnf Co-authored-by: marin <postma.marin@protonmail.com>	2021-11-09 01:00:42 +01:00
Tamo	9c36e497d9	Rename the key_component into a value_component Co-authored-by: marin <postma.marin@protonmail.com>	2021-11-09 00:59:44 +01:00
Irevoire	6515838d35	improve the readability of the _geoPoint thingy in the value	2021-11-09 00:57:46 +01:00
Tamo	ea52aff6dc	Rename the ExtendNomError trait to NomErrorExt Co-authored-by: marin <postma.marin@protonmail.com>	2021-11-09 00:52:17 +01:00
Irevoire	ef0d5a8240	flatten a match	2021-11-09 00:49:13 +01:00
Tamo	15bd14297e	Remove useless closure Co-authored-by: marin <postma.marin@protonmail.com>	2021-11-09 00:45:46 +01:00
Irevoire	21d115dcbb	remove greedy-error	2021-11-08 17:53:41 +01:00
Irevoire	959ca66125	improve the error diagnostic when parsing values	2021-11-08 15:58:21 +01:00
Tamo	7483c7513a	fix the filterable fields	2021-11-07 01:52:19 +01:00
Tamo	e5af3ac65c	rename the filter_condition.rs to filter.rs	2021-11-06 16:37:55 +01:00
Tamo	6831c23449	merge with main	2021-11-06 16:34:30 +01:00
Tamo	5c01e9bf7c	fix the benchmarks	2021-11-06 16:03:49 +01:00
Tamo	075d9c97c0	re-implement the equality between tokens to only compare the inner value	2021-11-06 16:02:27 +01:00
Tamo	b249989bef	fix most of the tests	2021-11-06 01:32:12 +01:00
Tamo	070ec9bd97	small update on the README	2021-11-05 17:45:20 +01:00
Tamo	27a6a26b4b	makes the parse function part of the filter_parser	2021-11-05 10:46:54 +01:00
Tamo	76d961cc77	implements the last errors	2021-11-04 17:42:06 +01:00
Tamo	8234f9fdf3	recreate most filter error except for the geosearch	2021-11-04 17:24:55 +01:00
Tamo	7328ffb034	stop panicking in case of internal error	2021-11-04 16:20:53 +01:00
Tamo	3e5550c910	clean the errors	2021-11-04 16:12:17 +01:00
Tamo	72a9071203	fix typo	2021-11-04 16:03:52 +01:00
Tamo	07a5ffb04c	update http-ui	2021-11-04 15:52:22 +01:00
Tamo	a58bc5bebb	update milli with the new parser_filter	2021-11-04 15:02:36 +01:00
Tamo	b1a0110a47	update the main	2021-11-04 14:48:39 +01:00
Tamo	d0fe9dea61	update the readme	2021-11-04 14:43:36 +01:00
Tamo	b165c77fa7	add a smol README	2021-11-04 14:39:02 +01:00
Tamo	54aec7ac5f	update the filter parser and some code for the fuzzer	2021-11-04 14:22:35 +01:00
bors[bot]	a2fc74f010	Merge #412 412: Change Attribute and Ranking rules errors r=ManyTheFish a=ManyTheFish # Pull Request Fixes Meilisearch [PR comment](https://github.com/meilisearch/MeiliSearch/pull/1873#issuecomment-959786406) Co-authored-by: many <maxime@meilisearch.com>	2021-11-04 13:08:50 +00:00
many	743ed9f57f	Bump milli version	2021-11-04 14:04:21 +01:00
many	7b3bac46a0	Change Attribute and Ranking rules errors	2021-11-04 13:19:32 +01:00
bors[bot]	3be37b00e7	Merge #410 410: Update version for the next release (v0.20.1) r=Kerollmops a=ManyTheFish Co-authored-by: many <maxime@meilisearch.com>	2021-11-03 13:32:03 +00:00
many	702589104d	Update version for the next release (v0.20.1)	2021-11-03 14:20:01 +01:00
bors[bot]	cb9e7e510b	Merge #408 408: Change last error messages r=ManyTheFish a=ManyTheFish Change forgotten error messages Co-authored-by: many <maxime@meilisearch.com>	2021-11-03 10:51:33 +00:00
many	0c0038488c	Change last error messages	2021-11-03 11:24:06 +01:00
Tamo	5d3af5f273	remove all genericity in favor of my custom error type	2021-11-02 20:27:07 +01:00
Tamo	76a2adb7c3	re-enable the tests in the parser and start the creation of an error type	2021-11-02 17:35:17 +01:00
bors[bot]	5a6d22d4ec	Merge #407 407: Update version for the next release (v0.20.0) r=curquiza a=curquiza Breaking because of #405 and #406 Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2021-10-28 13:43:48 +00:00
bors[bot]	08ae47e475	Merge #405 405: Change some error messages r=ManyTheFish a=ManyTheFish Co-authored-by: many <maxime@meilisearch.com>	2021-10-28 13:35:55 +00:00
Clémentine Urquizar	056ff13c4d	Update version for the next release (v0.20.0)	2021-10-28 14:52:57 +02:00
many	9f1e0d2a49	Refine asc/desc error messages	2021-10-28 14:47:17 +02:00
many	ed6db19681	Fix PR comments	2021-10-28 11:18:32 +02:00
bors[bot]	9875f2646a	Merge #406 406: return document count from builder r=MarinPostma a=MarinPostma `DocumentBatchBuilder::finish` now returns the number of documents in the batch. This is more compact that calling `len()` just before calling finish. Co-authored-by: marin postma <postma.marin@protonmail.com>	2021-10-28 08:42:38 +00:00
marin postma	183d3dada7	return document count from builder	2021-10-28 10:33:04 +02:00
many	2be755ce75	Lower error check, already check in meilisearch	2021-10-27 19:50:41 +02:00
many	3599df77f0	Change some error messages	2021-10-27 19:33:01 +02:00
bors[bot]	d7943fe225	Merge #402 402: Optimize document transform r=MarinPostma a=MarinPostma This pr optimizes the transform of documents additions in the obkv format. Instead on accepting any serializable objects, we instead treat json and CSV specifically: - For json, we build a serde `Visitor`, that transform the json straight into obkv without intermediate representation. - For csv, we directly write the lines in the obkv, applying other optimization as well. Co-authored-by: marin postma <postma.marin@protonmail.com>	2021-10-26 09:55:28 +00:00
bors[bot]	6758146213	Merge #404 404: remove search crate r=Kerollmops a=MarinPostma The functionalities of the search crate have been moved to the cli crate. The outstanding files are removed by this pr. Co-authored-by: marin postma <postma.marin@protonmail.com>	2021-10-26 09:40:34 +00:00
marin postma	9b8ab40d80	remove search folder	2021-10-26 11:35:49 +02:00
marin postma	baddd80069	implement review suggestions	2021-10-25 18:29:12 +02:00
marin postma	f9445c1d90	return float parsing error context in csv	2021-10-25 17:27:10 +02:00
bors[bot]	15c29cdd9b	Merge #401 401: Update version for the next release (v0.19.0) r=curquiza a=curquiza Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2021-10-25 12:49:53 +00:00
bors[bot]	13d8272173	Merge #403 403: Revert "Replacing pest with nom" r=curquiza a=curquiza Reverts meilisearch/milli#358 Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2021-10-25 12:16:49 +00:00
Clémentine Urquizar	208903ddde	Revert "Replacing pest with nom "	2021-10-25 11:58:00 +02:00
Clémentine Urquizar	679fe18b17	Update version for the next release (v0.19.0)	2021-10-25 11:52:17 +02:00
marin postma	3fcccc31b5	add document builder example	2021-10-25 10:26:43 +02:00
marin postma	430e9b13d3	add csv builder tests	2021-10-25 10:26:43 +02:00
marin postma	53c79e85f2	document errors	2021-10-25 10:26:43 +02:00
marin postma	2e62925a6e	fix tests	2021-10-25 10:26:42 +02:00
marin postma	0f86d6b28f	implement csv serialization	2021-10-25 10:26:42 +02:00
marin postma	8d70b01714	optimize document deserialization	2021-10-25 10:26:42 +02:00
Tamo	1327807caa	add some error messages	2021-10-22 19:00:33 +02:00
Tamo	c8d03046bf	add a check on the fid in the geosearch	2021-10-22 18:08:18 +02:00
Tamo	3942b3732f	re-implement the geosearch	2021-10-22 18:03:39 +02:00
Tamo	7cd9109e2f	lowercase value extracted from Token	2021-10-22 17:50:15 +02:00
Tamo	4e113bbf1b	handle the case of empty input	2021-10-22 17:49:08 +02:00
Tamo	e25ca9776f	start updating the exposed function to makes other modules happy	2021-10-22 17:23:22 +02:00
Tamo	6c9165b6a8	provide a helper to parse the token but to not handle the errors	2021-10-22 16:52:13 +02:00
Tamo	efb2f8b325	convert the errors	2021-10-22 16:38:35 +02:00
Tamo	d6ba84ea99	re introduce the special error type to be able to add context to the errors	2021-10-22 15:09:56 +02:00
Tamo	c27870e765	integrate a first version without any error handling	2021-10-22 14:33:18 +02:00
Tamo	01dedde1c9	update some names and move some parser out of the lib.rs	2021-10-22 01:59:38 +02:00
Tamo	7e5c5c4d27	start a new rewrite of the filter parser	2021-10-22 01:15:42 +02:00
Tamo	c634d43ac5	add a simple test on the filters with an integer	2021-10-21 17:10:27 +02:00
Tamo	6c15f50899	rewrite the parser logic	2021-10-21 16:45:42 +02:00
Tamo	e1d81342cf	add test on the or and and operator	2021-10-21 13:01:25 +02:00
Tamo	423baac08b	fix the tests	2021-10-21 12:45:40 +02:00
Tamo	36281a653f	write all the simple tests	2021-10-21 12:40:11 +02:00
Clémentine Urquizar	f8fe9316c0	Update version for the next release (v0.18.1)	2021-10-21 11:56:14 +02:00
Tamo	661bc21af5	Fix the filter parser And add a bunch of tests on the filter::from_array	2021-10-21 11:45:03 +02:00
bors[bot]	b6af84eb77	Merge #394 394: Added search_geo benchmark in cron job r=irevoire a=fumblehool fixes: #392 `search_geo` cron will run every friday at 18:30 Co-authored-by: Damanpreet Singh <daman.4880@gmail.com>	2021-10-18 14:33:32 +00:00
bors[bot]	7906461c14	Merge #396 396: Fix indexing benchmark GH actions upload filename r=irevoire a=fumblehool fixes: #393 Co-authored-by: Damanpreet Singh <daman.4880@gmail.com>	2021-10-18 13:34:10 +00:00
Damanpreet Singh	2e4604b0b9	fixed filename for search_* crons	2021-10-18 18:48:38 +05:30
Damanpreet Singh	4c34164d2e	fixed filename for search_geo cron	2021-10-18 18:43:36 +05:30
bors[bot]	9df4f3aaad	Merge #397 397: Fix typo in repo r=curquiza a=saintmalik Fix the single typo found in this repo Co-authored-by: SaintMalik <37118134+saintmalik@users.noreply.github.com>	2021-10-18 11:59:48 +00:00
bors[bot]	513d3178c6	Merge #398 398: Update version for the next release (v0.18.2) r=irevoire a=curquiza Breaking because of https://github.com/meilisearch/milli/pull/358 Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2021-10-18 11:47:26 +00:00
Clémentine Urquizar	2209acbfe2	Update version for the next release (v0.18.2)	2021-10-18 13:45:48 +02:00
SaintMalik	70121e3c6b	fix typo in repo	2021-10-18 04:00:19 +01:00
bors[bot]	59cc59e93e	Merge #358 358: Replacing pest with nom r=Kerollmops a=CNLHC Co-authored-by: 刘瀚骋 <cn_lhc@qq.com>	2021-10-16 20:44:38 +00:00
Damanpreet Singh	493d9b98f5	fix indexing benchmark GH actions upload filename	2021-10-16 21:52:36 +05:30
Damanpreet Singh	efaef4f748	Added search_geo benchmark in cron job	2021-10-16 21:41:45 +05:30
刘瀚骋	7666e4f34a	follow the suggestions	2021-10-14 21:37:59 +08:00
刘瀚骋	2ea2f7570c	use nightly cargo to format the code	2021-10-14 16:46:13 +08:00
刘瀚骋	e750465e15	check logic for geolocation.	2021-10-14 16:12:00 +08:00
bors[bot]	aa5e099718	Merge #390 390: Add helper methods on the settings r=Kerollmops a=irevoire This would be a good addition to look at the content of a setting without consuming it. It’s useful for analytics. Co-authored-by: Irevoire <tamo@meilisearch.com>	2021-10-13 20:36:30 +00:00
bors[bot]	c7db4176f3	Merge #384 384: Replace memmap with memmap2 r=Kerollmops a=palfrey [memmap is unmaintained](https://rustsec.org/advisories/RUSTSEC-2020-0077.html) and needs replacing. memmap2 is a drop-in replacement fork that's well maintained. Note that the version numbers got reset on fork, hence the lower values. Co-authored-by: Tom Parker-Shemilt <palfrey@tevp.net>	2021-10-13 13:47:23 +00:00
Irevoire	a3e7c468cd	add helper methods on the settings	2021-10-13 13:05:07 +02:00
刘瀚骋	cd359cd96e	WIP: extract the error trait bound to new trait.	2021-10-13 18:04:15 +08:00
刘瀚骋	5de5dd80a3	WIP: remove '_nom' suffix/redundant error enum/...	2021-10-13 11:06:15 +08:00
刘瀚骋	2c65781d91	format	2021-10-12 22:20:22 +08:00
bors[bot]	6e3b869e6a	Merge #388 388: fix primary key inference r=MarinPostma a=MarinPostma The primary key is was infered from a hashtable index of the field. For this reason the order in which the fields were interated upon was not deterministic, and the primary key was chosed ffrom the first field containing "id". This fix sorts the the index by field_id when infering the primary key. Co-authored-by: mpostma <postma.marin@protonmail.com>	2021-10-12 09:25:16 +00:00
mpostma	86ead92ed5	infer primary key on sorted fields	2021-10-12 11:15:11 +02:00
mpostma	9a266a531b	test correct primary key inference	2021-10-12 11:08:53 +02:00
bors[bot]	3f7f24b90e	Merge #368 368: Remove limit of 1000 position per attribute r=irevoire a=ManyTheFish Instead of using an arbitrary limit we encode the absolute position in a u32 using one strong u16 for the field id and a weak u16 for the relative position in the attribute. - [x] check database size difference below is the database size difference for each dataset: ![Capture d’écran 2021-09-27 à 18 01 44](https://user-images.githubusercontent.com/6482087/134944199-bd25fed0-6c34-475c-9afc-197871e06553.png) - [ ] check search time on big dataset Related to [product#202](https://github.com/meilisearch/product/issues/202) Co-authored-by: many <maxime@meilisearch.com>	2021-10-12 08:30:33 +00:00
many	c5a6075484	Make max_position_per_attributes changable	2021-10-12 10:10:50 +02:00
many	360c5ff3df	Remove limit of 1000 position per attribute Instead of using an arbitrary limit we encode the absolute position in a u32 using one strong u16 for the field id and a weak u16 for the relative position in the attribute.	2021-10-12 10:10:50 +02:00
刘瀚骋	d323e35001	add a test case	2021-10-12 13:30:40 +08:00
刘瀚骋	70f576d5d3	error handling	2021-10-12 13:30:40 +08:00
刘瀚骋	28f9be8d7c	support syntax	2021-10-12 13:30:40 +08:00
刘瀚骋	469d92c569	tweak error handling	2021-10-12 13:30:40 +08:00
刘瀚骋	7a90a101ee	reorganize parser logic	2021-10-12 13:30:40 +08:00
刘瀚骋	f7796edc7e	remove everything about pest	2021-10-12 13:30:40 +08:00
刘瀚骋	ac1df9d9d7	fix typo and remove pest	2021-10-12 13:30:40 +08:00
刘瀚骋	50ad750ec1	enhance error handling	2021-10-12 13:30:40 +08:00
刘瀚骋	8748df2ca4	draft without error handling	2021-10-12 13:30:40 +08:00
bors[bot]	8f6b6c9042	Merge #385 385: Fix the wiki indexing benchmark r=ManyTheFish a=irevoire Co-authored-by: Tamo <tamo@meilisearch.com>	2021-10-11 15:12:24 +00:00
bors[bot]	07fb6d64e5	Merge #386 386: fix obkv document r=curquiza a=MarinPostma When serializing a document, the serializer resolved the field_id of the current field and immediately added it to the obkv document under construction. The issue with that is that obkv expects the fields to be inserted in order, and when a document with out of order fields was added, obkv failed to insert the field. The current fix first resolves each field_id, and adds all the fields to a temporary `BTreeMap`, until `end` is called on the map serializer, where all the fields are added to the obkv at once, and in order. Co-authored-by: mpostma <postma.marin@protonmail.com>	2021-10-11 13:45:04 +00:00
bors[bot]	e45c846af5	Merge #387 387: Update version for the next release (v0.17.2) r=Kerollmops a=curquiza Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2021-10-11 13:21:47 +00:00
Clémentine Urquizar	dd56e82dba	Update version for the next release (v0.17.2)	2021-10-11 15:20:35 +02:00
mpostma	99889a0ed0	add obkv document serialization test	2021-10-11 15:13:17 +02:00
mpostma	799f3d43c8	fix serialization to obkv format	2021-10-11 15:04:47 +02:00
Tamo	ed7fd855af	fix the wiki indexing benchmark	2021-10-11 14:26:36 +02:00
Tom Parker-Shemilt	2dfe24f067	memmap -> memmap2	2021-10-10 22:47:12 +01:00
bors[bot]	a2743baaa3	Merge #383 383: Add check on latitude and longitude r=irevoire a=irevoire Latitudes are not supposed to go beyond 90 degrees or below -90. The same goes for longitudes with 180 or -180. This was badly implemented in the filters, and was not implemented for the `AscDesc` rules. Co-authored-by: Tamo <tamo@meilisearch.com> Co-authored-by: Irevoire <tamo@meilisearch.com>	2021-10-08 10:15:25 +00:00
Irevoire	b65aa7b5ac	Apply suggestions from code review Co-authored-by: Clément Renault <clement@meilisearch.com>	2021-10-07 17:51:52 +02:00
Tamo	11dfe38761	Update the check on the latitude and longitude Latitude are not supposed to go beyound 90 degrees or below -90. The same goes for longitude with 180 or -180. This was badly implemented in the filters, and was not implemented for the AscDesc rules.	2021-10-07 16:10:43 +02:00
bors[bot]	dde1da1c0e	Merge #382 382: Refactor attribute criterion r=Kerollmops a=ManyTheFish ### Re-implement set based algorithm for attribute criterion #### Levels Instead of doing level iteration and digging in the interesting level, we only iterate over the lowest level. #### crossword iteration VS minimal position iteration Instead of crossing word position in order to iterate strictly over the position that gives the best rank in good order; we iterate word by word starting with the word that increases the rank the little as possible. This new method is a bit less precise but way simpler. ### Simplify word-level-position database We don't use levels anymore in the attribute criterion, and so we removed the level complexity of the database making a word-position-docids database. ### Benchmarks on search on big datasets #### songs main VS refactor-attribute-criterion ```diff group search_songsmain_31c18f09 search_songsrefactor-attribute-criterion_1bd15d84 ----- ------------------------- ------------------------------------------------- - smol-songs.csv: basic filter: <=/Notstandskomitee 1.00 84.8±0.58µs ? ?/sec 1.09 92.2±8.98µs ? ?/sec + smol-songs.csv: basic filter: TO/Notstandskomitee 1.18 98.0±6.30µs ? ?/sec 1.00 83.2±0.97µs ? ?/sec + smol-songs.csv: basic with quote/"david" "bowie" 114.68 76.0±0.20ms ? ?/sec 1.00 662.5±5.03µs ? ?/sec - smol-songs.csv: basic with quote/"john" 1.00 197.4±1.06µs ? ?/sec 1.05 208.1±1.53µs ? ?/sec + smol-songs.csv: basic with quote/"michael" "jackson" 2.75 2.0±0.01ms ? ?/sec 1.00 738.9±3.91µs ? ?/sec + smol-songs.csv: basic without quote/david bowie 297.42 1499.3±0.86ms ? ?/sec 1.00 5.0±0.02ms ? ?/sec + smol-songs.csv: basic without quote/michael jackson 2.55 8.9±0.02ms ? ?/sec 1.00 3.5±0.01ms ? ?/sec + smol-songs.csv: big filter/john 1.08 473.6±2.25µs ? ?/sec 1.00 438.1±2.59µs ? ?/sec - smol-songs.csv: prefix search/a 1.00 446.9±1.81µs ? ?/sec 1.79 800.5±4.45µs ? ?/sec - smol-songs.csv: prefix search/b 1.00 398.5±2.74µs ? ?/sec 1.81 723.1±5.46µs ? ?/sec - smol-songs.csv: prefix search/i 1.00 486.3±1.99µs ? ?/sec 1.69 823.6±9.42µs ? ?/sec - smol-songs.csv: prefix search/s 1.00 229.6±3.29µs ? ?/sec 2.59 594.4±2.22µs ? ?/sec - smol-songs.csv: prefix search/x 1.00 150.2±0.76µs ? ?/sec 1.11 166.0±0.87µs ? ?/sec ``` On songs, the new algorithm gives a big improvement on slow queries, and is slower on one char prefix search (fast queries <1ms). #### wiki main VS refactor-attribute-criterion ```diff group search_wikimain_31c18f09 search_wikirefactor-attribute-criterion_1bd15d84 ----- ------------------------ ------------------------------------------------ - smol-wiki-articles.csv: basic with quote/"rock" "and" "roll" 1.00 3.2±0.01ms ? ?/sec 1.15 3.7±0.01ms ? ?/sec - smol-wiki-articles.csv: basic without quote/film 1.00 351.5±2.47µs ? ?/sec 1.13 396.8±1.63µs ? ?/sec + smol-wiki-articles.csv: basic without quote/rock and roll 1.10 9.4±0.02ms ? ?/sec 1.00 8.6±0.04ms ? ?/sec - smol-wiki-articles.csv: basic without quote/spain 1.00 446.0±3.23µs ? ?/sec 1.11 496.6±7.75µs ? ?/sec - smol-wiki-articles.csv: prefix search/c 1.00 115.6±0.61µs ? ?/sec 2.22 256.7±1.24µs ? ?/sec - smol-wiki-articles.csv: prefix search/g 1.00 189.7±2.03µs ? ?/sec 1.57 297.0±1.35µs ? ?/sec - smol-wiki-articles.csv: prefix search/j 1.00 209.2±1.11µs ? ?/sec 1.40 293.0±2.09µs ? ?/sec - smol-wiki-articles.csv: prefix search/q 1.00 79.0±0.44µs ? ?/sec 1.10 87.2±0.69µs ? ?/sec - smol-wiki-articles.csv: prefix search/t 1.00 270.1±1.15µs ? ?/sec 1.55 419.9±5.16µs ? ?/sec - smol-wiki-articles.csv: prefix search/x 1.00 244.9±1.33µs ? ?/sec 1.07 260.9±1.95µs ? ?/sec - smol-wiki-articles.csv: words/Abraham machin 1.00 8.1±0.03ms ? ?/sec 1.17 9.4±0.02ms ? ?/sec - smol-wiki-articles.csv: words/Idaho Bellevue pizza 1.00 19.3±0.07ms ? ?/sec 1.07 20.6±0.05ms ? ?/sec ``` On wiki we have some regressions `+17%` and `+15%` on request `>1ms`. Co-authored-by: many <maxime@meilisearch.com>	2021-10-06 09:19:33 +00:00
many	085bc6440c	Apply PR comments	2021-10-06 11:12:26 +02:00
many	1bd15d849b	Reduce candidates threshold	2021-10-05 18:52:14 +02:00
many	ea4bd29d14	Apply PR comments	2021-10-05 17:35:07 +02:00
many	5ed75de0db	Update infos crate	2021-10-05 13:56:12 +02:00
many	3296bb243c	Simplify word level position DB into a word position DB	2021-10-05 12:15:02 +02:00
many	75d341d928	Re-implement set based algorithm for attribute criterion	2021-10-05 12:14:50 +02:00
bors[bot]	31c18f0953	Merge #381 381: Update version for the next release (v0.17.1) r=irevoire a=curquiza Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2021-10-03 02:12:43 +00:00
Clémentine Urquizar	05d8a33a28	Update version for the next release (v0.17.1)	2021-10-02 16:21:31 +02:00
bors[bot]	c9092c72bf	Merge #380 380: Reserved keyword error message r=Kerollmops a=irevoire And I missed _another_ reserved keyword error message in the filter :( Co-authored-by: Tamo <tamo@meilisearch.com>	2021-10-01 07:13:31 +00:00
Tamo	d9eba9d145	improve and test the sort error message	2021-09-30 14:38:27 +02:00
Tamo	0ee67bb7d1	improve the reserved keyword error message for the filters	2021-09-30 14:38:27 +02:00
bors[bot]	22551d0941	Merge #379 379: Revert "Change chunk size to 4MiB to fit more the end user usage" r=curquiza a=ManyTheFish Reverts meilisearch/milli#370 Co-authored-by: Many <legendre.maxime.isn@gmail.com>	2021-09-29 13:20:53 +00:00
Many	26b5dad042	Revert "Change chunk size to 4MiB to fit more the end user usage"	2021-09-29 15:08:39 +02:00
bors[bot]	6a057a3bd0	Merge #378 378: Hotfix meilisearch#1707 r=Kerollmops a=ManyTheFish This PR contains an ugly quick fix of [meilisearch#1707](https://github.com/meilisearch/MeiliSearch/issues/1707). - remove comparison reverse on rank. Enhancing relevancy and performances - iterate over level 0 only. Enhancing performances. A better fix is in development. Co-authored-by: many <maxime@meilisearch.com> Co-authored-by: Many <legendre.maxime.isn@gmail.com>	2021-09-29 12:57:31 +00:00
Many	2e49230ca2	Update milli/src/search/criteria/attribute.rs Co-authored-by: Clément Renault <clement@meilisearch.com>	2021-09-29 14:49:45 +02:00
Many	7ad0214089	Update milli/src/search/criteria/attribute.rs Co-authored-by: Clément Renault <clement@meilisearch.com>	2021-09-29 14:49:41 +02:00
many	1df5b8712b	Hotfix meilisearch#1707	2021-09-29 14:41:56 +02:00
bors[bot]	bfedbc1b6d	Merge #374 374: Enhance CSV document parsing r=Kerollmops a=ManyTheFish Benchmarks on `search_songs` were crashing because of the CSV parsing. Co-authored-by: many <maxime@meilisearch.com>	2021-09-29 08:55:54 +00:00
bors[bot]	68c758a533	Merge #376 376: Stop casting integer docids to string r=Kerollmops a=irevoire When a docid is an integer, we stop casting it to a string, and thus we don't add `"` around it. Co-authored-by: Tamo <tamo@meilisearch.com>	2021-09-29 08:32:48 +00:00
many	d2427f18e5	Enhance CSV document parsing	2021-09-29 10:25:33 +02:00
bors[bot]	00f94b1ffd	Merge #377 377: Update version for the next release (v0.17.0) r=Kerollmops a=curquiza Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2021-09-28 20:43:33 +00:00
Clémentine Urquizar	0e8665bf18	Update version for the next release (v0.17.0)	2021-09-28 19:38:12 +02:00
Tamo	f65153ad64	stop casting integer docids to string	2021-09-28 18:35:54 +02:00
bors[bot]	adddf3f179	Merge #375 375: Fixes #365 r=Kerollmops a=vishnugt Co-authored-by: Vishnu Ganesan <vganesan@microsoft.com> Co-authored-by: Vishnu Gt <vishnugt@hotmail.com>	2021-09-28 14:42:48 +00:00
Vishnu Gt	785c1372f2	Change "settings" to "setting" Co-authored-by: Clément Renault <renault.cle@gmail.com>	2021-09-28 20:11:32 +05:30
Vishnu Ganesan	3580b2d803	Fixes #365	2021-09-28 19:30:23 +05:30
bors[bot]	3a12f5887e	Merge #373 373: Improve error message for bad sort syntax with geosearch r=Kerollmops a=irevoire `@Kerollmops` This should be the last PR for the geosearch and error handling, sorry for doing it in so many steps 😬 Co-authored-by: Tamo <tamo@meilisearch.com>	2021-09-28 12:39:32 +00:00
Tamo	a80dcfd4a3	improve error message for bad sort syntax with geosearch	2021-09-28 14:32:24 +02:00
bors[bot]	b2a332599e	Merge #372 372: Fix Meilisearch 1714 r=Kerollmops a=ManyTheFish The bug comes from the typo tolerance, to know how many typos are accepted we were counting bytes instead of characters in a word. On Chinese Script characters, we were allowing 2 typos on 3 characters words. We are now counting the number of char instead of counting bytes to assign the typo tolerance. Related to [Meilisearch#1714](https://github.com/meilisearch/MeiliSearch/issues/1714) Co-authored-by: many <maxime@meilisearch.com>	2021-09-28 11:59:45 +00:00
many	8046ae4bd5	Count the number of char instead of counting bytes to assign the typo tolerance	2021-09-28 12:10:43 +02:00
many	1988416295	Add failing test related to Meilisearch#1714	2021-09-28 12:05:11 +02:00
bors[bot]	3b479948c6	Merge #371 371: Provide a sort error handler r=Kerollmops a=irevoire This PR simplify the error handling of asc-desc rules for Meilisearch or any other wrapper by providing directly in milli a new error type called `SortError` that can be generated from an `AscDescError` and that can be automatically converted to a `UserError`. Basically now, wherever you are in the code as a user or in milli you can parse an `AscDesc` syntax and depending on the context, cast it either as a `SortError` or a `CriterionError` in one line with improved error messages. Co-authored-by: Tamo <tamo@meilisearch.com>	2021-09-28 09:28:32 +00:00
Tamo	cc732fe95e	update http-ui to use the sort-error	2021-09-28 11:15:24 +02:00
Tamo	c7cb816ae1	simplify the error handling of the sort syntax for meilisearch	2021-09-27 19:07:22 +02:00
bors[bot]	4c09f6838f	Merge #370 370: Change chunk size to 4MiB to fit more the end user usage r=ManyTheFish a=ManyTheFish We made several indexing tests using different sizes of datasets (5 datasets from 9MiB to 100MiB) on several typologies of VMs (`XS: 1GiB RAM, 1 VCPU`, `S: 2GiB RAM, 2 VCPU`, `M: 4GiB RAM, 3 VCPU`, `L: 8GiB RAM, 4 VCPU`). The result of these tests shows that the `4MiB` chunk size seems to be the best size compared to other chunk sizes (`2Mib`, `4MiB`, `8Mib`, `16Mib`, `32Mib`, `64Mib`, `128Mib`). below is the average time per chunk size: ![Capture d’écran 2021-09-27 à 14 27 50](https://user-images.githubusercontent.com/6482087/134909368-ef0bc45e-68d5-49d1-aaf9-91113b7c410f.png) <details> <summary>Detailled data</summary> <br> ![Capture d’écran 2021-09-27 à 14 39 48](https://user-images.githubusercontent.com/6482087/134909952-a36b1457-bbbd-4a6c-bbe5-519e4b926b5a.png) </br> </details> Co-authored-by: many <maxime@meilisearch.com>	2021-09-27 12:57:52 +00:00
many	b188063869	Change chunk size to 4MiB to fit more the end user usage	2021-09-27 14:26:21 +02:00
bors[bot]	0f8320bdc2	Merge #369 369: Add test checking the bug reported in meilisearch issue 1716 r=Kerollmops a=ManyTheFish The bug is not present in the newer milli version. Related to [Meilisearch#1716](https://github.com/meilisearch/MeiliSearch/issues/1716) Co-authored-by: many <maxime@meilisearch.com>	2021-09-23 14:27:34 +00:00
many	551df0cb77	Add test checking the bug reported in meilisearch issue 1716	2021-09-23 15:55:39 +02:00
bors[bot]	87dd441a3a	Merge #367 367: Update version for the next release (v0.16.0) r=Kerollmops a=curquiza Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2021-09-22 15:20:20 +00:00
Clémentine Urquizar	1eacab2169	Update version for the next release (v0.15.1)	2021-09-22 17:18:54 +02:00
bors[bot]	b806097141	Merge #366 366: Geosearch error handling r=Kerollmops a=irevoire Rewrite most of geosearch error handling and another batch of tests on the criterion parsing. Co-authored-by: Tamo <tamo@meilisearch.com> Co-authored-by: Irevoire <tamo@meilisearch.com>	2021-09-22 15:08:11 +00:00
Irevoire	218f0a6661	Apply suggestions from code review Co-authored-by: Clément Renault <clement@meilisearch.com>	2021-09-22 17:00:27 +02:00
Tamo	47ee93b0bd	return an error when _geoPoint is used but _geo is not sortable	2021-09-22 16:37:41 +02:00
Tamo	1e5e3d57e2	auto convert AscDescError into CriterionError	2021-09-22 16:37:41 +02:00
Tamo	023446ecf3	create a smaller and easier to maintain CriterionError type	2021-09-22 16:37:41 +02:00
Tamo	86e272856a	create an asc_desc error type that is never supposed to be returned to the end user	2021-09-22 16:37:41 +02:00
Tamo	257e621d40	create an asc_desc module	2021-09-22 16:37:41 +02:00
Tamo	113a061bee	fix the error handling on the criterion side	2021-09-22 15:09:07 +02:00
bors[bot]	ad3befaaf5	Merge #364 364: Fix all the benchmarks r=Kerollmops a=irevoire #324 broke all benchmarks. I fixed everything and noticed that `cargo check --all` was insufficient to check the bench in multiple workspaces, so I also updated the CI to use `cargo check --workspace --all-targets`. Co-authored-by: Tamo <tamo@meilisearch.com>	2021-09-22 12:40:34 +00:00
Tamo	176160d32f	fix all benchmarks and add the compile time checking of the benhcmarks in the ci	2021-09-22 12:10:21 +02:00
bors[bot]	16790ee620	Merge #363 363: Fix the returned `AscDesc` error r=Kerollmops a=irevoire With my previous PR on the geosearch I erased the change I've introduced with my pre-previous PR about the new error type when we fail to parse the `AscDesc` type. Sorry for that, here is the fix Co-authored-by: Tamo <tamo@meilisearch.com>	2021-09-22 09:53:35 +00:00
Tamo	78b0bce9a1	fix the returned error when asc desc fails to be parsed	2021-09-22 11:37:05 +02:00
bors[bot]	2837cab5da	Merge #362 362: Remove the `Cargo.lock` again r=Kerollmops a=irevoire Co-authored-by: Tamo <tamo@meilisearch.com>	2021-09-22 09:33:09 +00:00
Tamo	2e99fa8251	remove the cargo.lock again	2021-09-22 11:30:33 +02:00
bors[bot]	fe9f380993	Merge #361 361: Update version for the next release (v0.15.0) r=Kerollmops a=curquiza Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2021-09-21 16:19:16 +00:00
Clémentine Urquizar	f8ecbc28e2	Update version for the next release (v0.15.0)	2021-09-21 18:09:14 +02:00
bors[bot]	700318dc62	Merge #357 357: Add benchmarks for the geosearch r=Kerollmops a=irevoire closes #336 Should I merge this PR in #322 and then we merge everything in `main` or should we wait for #322 to be merged and then merge this one in `main` later? Co-authored-by: Tamo <tamo@meilisearch.com> Co-authored-by: Irevoire <tamo@meilisearch.com>	2021-09-21 16:08:06 +00:00
bors[bot]	9d9010e45f	Merge #324 324: Implement documents API r=Kerollmops a=MarinPostma This pr implement the intermediary document representation for milli. The JSON, JSONL and CSV formats are replaced with the format instead, to push the serialization duty on the client side. The `documents` module contains the interface to the new document format: - The `DocumentsBuilder` allows the creation of a writer backed document addition, when documents are added either one by one, or as arrays of depth 1. This is made possible by the fact that the seriliazer used by the `add_documents` methods only accepts `[Object]` and `Object`. The related serialization logic is located in the `serde.rs` file. - The `DocumentsReader` allows to to iterate over the documents created by a `DocumentsBuilder`. A call to `next_document_with_index` returns the next obkv reader in the document addition, along with a reference to the index used to map the field ids in the obkv reader to the field names All references to json, jsonl or csv in the tests have been replaced with the `documents!` macro, works exaclty like the `serde_json::json` macro, as a convenient way to create a `DocumentsReader`. Rewrote the search cli, to the `cli` crate, to also allow index manipulation. This only offers basic functionalities for now, but is meant to be easier to extend than http ui blocked by #308 Co-authored-by: mpostma <postma.marin@protonmail.com>	2021-09-21 15:40:03 +00:00
mpostma	aa6c5df0bc	Implement documents format document reader transform remove update format support document sequences fix document transform clean transform improve error handling add documents! macro fix transform bug fix tests remove csv dependency Add comments on the transform process replace search cli fmt review edits fix http ui fix clippy warnings Revert "fix clippy warnings" This reverts commit a1ce3cd96e603633dbf43e9e0b12b2453c9c5620. fix review comments remove smallvec in transform loop review edits	2021-09-21 16:58:33 +02:00
bors[bot]	94764e5c7c	Merge #360 360: Update version for the next release (v0.14.0) r=Kerollmops a=curquiza Release containing the geosearch, cf #322 Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2021-09-21 08:43:27 +00:00
bors[bot]	31c8de1cca	Merge #322 322: Geosearch r=ManyTheFish a=irevoire This PR introduces [basic geo-search functionalities](https://github.com/meilisearch/specifications/pull/59), it makes the engine able to index, filter and, sort by geo-point. We decided to use [the rstar library](https://docs.rs/rstar) and to save the points in [an RTree](https://docs.rs/rstar/0.9.1/rstar/struct.RTree.html) that we de/serialize in the index database [by using serde](https://serde.rs/) with [bincode](https://docs.rs/bincode). This is not an efficient way to query this tree as it will consume a lot of CPU and memory when a search is made, but at least it is an easy first way to do so. ### What we will have to do on the indexing part: - [x] Index the `_geo` fields from the documents. - [x] Create a new module with an extractor in the `extract` module that takes the `obkv_documents` and retrieves the latitude and longitude coordinates, outputting them in a `grenad::Reader` for further process. - [x] Call the extractor in the `extract::extract_documents_data` function and send the result to the `TypedChunk` module. - [x] Get the `grenad::Reader` in the `typed_chunk::write_typed_chunk_into_index` function and store all the points in the `rtree` - [x] Delete the documents from the `RTree` when deleting documents from the database. All this can be done in the `delete_documents.rs` file by getting the data structure and removing the points from it, inserting it back after the modification. - [x] Clearing the `RTree` entirely when we clear the documents from the database, everything happens in the `clear_documents.rs` file. - [x] save a Roaring bitmap of all documents containing the `_geo` field ### What we will have to do on the query part: - [x] Filter the documents at a certain distance around a point, this is done by [collecting the documents from the searched point](https://docs.rs/rstar/0.9.1/rstar/struct.RTree.html#method.nearest_neighbor_iter) while they are in range. - [x] We must introduce new `geoLowerThan` and `geoGreaterThan` variants to the `Operator` filter enum. - [x] Implement the `negative` method on both variants where the `geoGreaterThan` variant is implemented by executing the `geoLowerThan` and removing the results found from the whole list of geo faceted documents. - [x] Add the `_geoRadius` function in the pest parser. - [x] Introduce a `_geo` ascending ranking function that takes a point in parameter, ~~this function must keep the iterator on the `RTree` and make it peekable~~ This was not possible for now, we had to collect the whole iterator. Only the documents that are part of the candidates must be sent too! - [x] This ascending ranking rule will only be active if the search is set up with the `_geoPoint` parameter that indicates the center point of the ascending ranking rule. ----------- - On Meilisearch part: We must introduce a new concept, returning the documents with a new `_geoDistance` field when it passed by the `_geo` ranking rule, this has never been done before. We could maybe just do it afterward when the documents have been retrieved from the database, computing the distance from the `_geoPoint` and all of the documents to be returned. Co-authored-by: Irevoire <tamo@meilisearch.com> Co-authored-by: cvermand <33010418+bidoubiwa@users.noreply.github.com> Co-authored-by: Tamo <tamo@meilisearch.com>	2021-09-20 19:04:57 +00:00
Irevoire	0d104a0fce	Update milli/src/criterion.rs Co-authored-by: Clément Renault <clement@meilisearch.com>	2021-09-20 18:13:17 +02:00
Clémentine Urquizar	3f1453f470	Update version for the next release (v0.14.0)	2021-09-20 18:12:23 +02:00
Tamo	f4b8e5675d	move the reserved keyword logic for the criterion and sort + add test	2021-09-20 17:21:02 +02:00
Irevoire	3b7a2cdbce	fix typo Co-authored-by: Clément Renault <clement@meilisearch.com>	2021-09-20 16:10:39 +02:00
bors[bot]	203aa727a7	Merge #359 359: Improve the benchmark comparison script r=irevoire a=irevoire This modification allow us to compare more than 2 benchmarks or to only print the results of one benchmark Co-authored-by: Irevoire <tamo@meilisearch.com> Co-authored-by: Tamo <tamo@meilisearch.com>	2021-09-20 12:39:59 +00:00
Tamo	eaba772f21	update the README to better match the new critcmp usage Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2021-09-20 10:59:55 +02:00
Irevoire	9a920d1f93	Fix datasets links in the readme Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2021-09-20 10:44:37 +02:00
Tamo	5e683ba472	add benchmarks for the geosearch	2021-09-20 10:44:37 +02:00
Irevoire	f6c6b026bb	improve the comparison script	2021-09-16 11:25:51 +02:00
Tamo	c695a1ffd2	add the possibility to sort by descending order on geoPoint	2021-09-15 11:49:58 +02:00
Tamo	91ce4d1721	Stop iterating through the whole list of points We stop when there is no possible candidates left	2021-09-15 11:49:58 +02:00
bors[bot]	3b1885859d	Merge #356 356: Update the README r=curquiza a=Kerollmops This PR updates a little bit the README and more specifically the indexing times, fixes #352. Co-authored-by: Kerollmops <clement@meilisearch.com>	2021-09-14 10:13:05 +00:00
Kerollmops	2741aa8589	Update the indexing timings in the README	2021-09-14 11:42:59 +02:00
Kerollmops	a43f99c600	Inform the users that documents must have an id in there documents	2021-09-13 14:01:02 +02:00
bors[bot]	90d64d257f	Merge #354 354: Update version for the next release (v0.13.1) r=ManyTheFish a=curquiza Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2021-09-13 09:30:07 +00:00
Clémentine Urquizar	f167f7b412	Update version for the next release (v0.13.1)	2021-09-10 09:48:17 +02:00
bors[bot]	4af31ec9a6	Merge #353 353: Add lacking parameter to word level position builder r=Kerollmops a=ManyTheFish Co-authored-by: many <maxime@meilisearch.com>	2021-09-09 16:36:33 +00:00
Tamo	cfc62a1c15	use geoutils instead of haversine	2021-09-09 18:11:38 +02:00
many	26deeb45a3	Add lacking parameter to word level position builder	2021-09-09 17:49:04 +02:00
Tamo	3fc145c254	if we have no rtree we return all other provided documents	2021-09-09 17:44:09 +02:00
Irevoire	a84f3a8b31	Apply suggestions from code review Co-authored-by: Clément Renault <clement@meilisearch.com>	2021-09-09 15:09:35 +02:00
Tamo	c81ff22c5b	delete the invalid criterion name error in favor of invalid ranking rule name	2021-09-08 19:17:00 +02:00
Tamo	bad8ea47d5	edit the two lasts TODO comments	2021-09-08 18:24:09 +02:00
Tamo	b15c77ebc4	return an error in case a user try to sort with :desc	2021-09-08 18:24:09 +02:00
Tamo	4b618b95e4	rebase on main	2021-09-08 18:24:09 +02:00
Tamo	2988d3c76d	tests the geo filters	2021-09-08 18:24:09 +02:00
Tamo	e5ef0cad9a	use meters in the filters	2021-09-08 18:24:09 +02:00
Tamo	4f69b190bc	remove the distance from the search, the computation of the distance will be made on meilisearch side	2021-09-08 18:24:09 +02:00
Tamo	7ae2a7341c	introduce the reserved keywords in the filters	2021-09-08 18:24:09 +02:00
Tamo	6d5762a6c8	handle the case where you forgot entirely the parenthesis	2021-09-08 18:24:09 +02:00
Tamo	ebf82ac28c	improve the error messages and add tests for the filters	2021-09-08 18:24:09 +02:00
Tamo	bd4c248292	improve the error handling in general and introduce the concept of reserved keywords	2021-09-08 18:24:09 +02:00
Tamo	e8c093c1d0	fix the error handling in the filters	2021-09-08 18:24:09 +02:00
Tamo	f0b74637dc	fix all the tests	2021-09-08 18:24:09 +02:00
Tamo	b1bf7d4f40	reformat	2021-09-08 18:24:09 +02:00
Tamo	aca707413c	remove the memory leak	2021-09-08 18:24:09 +02:00
Tamo	a8a1f5bd55	move the geosearch criteria out of asc_desc.rs	2021-09-08 18:24:09 +02:00
Tamo	dc84ecc40b	fix a bug	2021-09-08 18:24:09 +02:00
Tamo	7483614b75	[HTTP-UI] add the sorters	2021-09-08 18:24:09 +02:00
Tamo	4820ac71a6	allow spaces in a geoRadius	2021-09-08 18:24:09 +02:00
Tamo	13c78e5aa2	Implement the _geoPoint in the sortable	2021-09-08 18:24:09 +02:00
Tamo	5bb175fc90	only index _geo if it's set as sortable OR filterable and only allow the filters if geo was set to filterable	2021-09-08 17:51:08 +02:00
Tamo	f73273d71c	only call the extractor if needed	2021-09-08 17:51:08 +02:00
cvermand	4fd0116a0d	Stringify objects on dashboard to avoid [Object object]	2021-09-08 17:51:08 +02:00
Irevoire	ea2f2ecf96	create a new database containing all the documents that were geo-faceted	2021-09-08 17:51:08 +02:00
Irevoire	4b459768a0	create the _geoRadius filter	2021-09-08 17:51:07 +02:00
Irevoire	6d70978edc	update the facet filter grammar	2021-09-08 17:51:07 +02:00
Irevoire	216a8aa3b2	add a tests for the indexation of the geosearch	2021-09-08 17:51:07 +02:00
Irevoire	a21c854790	handle errors	2021-09-08 17:51:07 +02:00
Irevoire	70ab2c37c5	remove multiple bugs	2021-09-08 17:51:07 +02:00
Irevoire	b4b6ba6d82	rename all the ’long’ into ’lng’ like written in the specification	2021-09-08 17:51:07 +02:00
Irevoire	3b9f1db061	implement the clear of the rtree	2021-09-08 17:51:07 +02:00
Irevoire	d344489c12	implement the deletion of geo points	2021-09-08 17:51:07 +02:00
Irevoire	44d6b6ae9e	Index the geo points	2021-09-08 17:51:07 +02:00
Irevoire	8d9c2c4425	create a new db with getters and setters	2021-09-08 17:51:07 +02:00
bors[bot]	b22aac92ac	Merge #342 342: Let the caller decide what kind of error they want to returns when parsing `AscDesc` r=Kerollmops a=irevoire This is one possible fix for #339 We would then need to patch these lines https://github.com/meilisearch/MeiliSearch/blob/main/meilisearch-http/src/index/search.rs#L110-L114 to return the error we want. Another solution would be to add a parameter to the `from_str` to specify which context we are in. Co-authored-by: Tamo <tamo@meilisearch.com>	2021-09-08 14:18:57 +00:00
Tamo	932998f5cc	let the caller decide if they want to return an invalidSortName or an invalidCriterionName error	2021-09-08 16:17:31 +02:00
bors[bot]	86c3b0c8c2	Merge #350 350: Fix mdb val size error r=Kerollmops a=ManyTheFish Related to [#1677](https://github.com/meilisearch/MeiliSearch/issues/1677) Co-authored-by: many <maxime@meilisearch.com>	2021-09-08 13:32:15 +00:00
many	e54280fbfc	Skip empty normalized words	2021-09-08 15:25:23 +02:00
many	d18ee58ab9	Check if key are not empty in validator	2021-09-08 15:25:23 +02:00
bors[bot]	63bc231243	Merge #349 349: Enable the grenad tempfile feature back r=ManyTheFish a=Kerollmops This PR enables the grenad `tempfile` feature back, [when this is feature is disabled the sorter writes the entries in memory](`7c082d05bf/src/sorter.rs (L470-L476)`) instead of on disk and therefore, consumes more memory. By enabling this feature grenad merges on disk by using the `tempfile` dependency. This PR also bumps milli to v0.3.1 where `@ManyTheFish` added an assert for when the allocator can't allocate and disable the default snappy compression in the `http-ui` crate. Co-authored-by: Kerollmops <clement@meilisearch.com>	2021-09-08 13:23:57 +00:00
Kerollmops	68856e5e2f	Disable the default snappy compression for the http-ui crate	2021-09-08 14:17:32 +02:00
Kerollmops	8a088fb99e	Bump grenad to v0.3.1	2021-09-08 14:08:55 +02:00
Kerollmops	20ad43b908	Enable the grenad tempfile feature back	2021-09-08 14:06:28 +02:00
bors[bot]	772e55d174	Merge #347 347: Update version for the next release (v0.13.0) r=curquiza a=curquiza Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2021-09-08 11:41:15 +00:00
bors[bot]	d160305868	Merge #348 348: Drop sorter before creating a new one r=Kerollmops a=ManyTheFish Co-authored-by: many <maxime@meilisearch.com>	2021-09-08 11:34:20 +00:00
many	9961b78b06	Drop sorter before creating a new one	2021-09-08 13:30:26 +02:00
Clémentine Urquizar	eb7b9d9dbf	Update version for the next release (v0.13.0)	2021-09-08 10:59:30 +02:00
bors[bot]	f5e418ace7	Merge #345 345: Better dependencies cache for CI r=irevoire a=shekhirin Co-authored-by: Alexey Shekhirin <a.shekhirin@gmail.com>	2021-09-08 08:43:19 +00:00
bors[bot]	48d211b8b0	Merge #344 344: Move the sort ranking rule before the exactness ranking rule r=ManyTheFish a=Kerollmops This PR moves the sort ranking rule at the 5th position by default, right before the exactness one. Co-authored-by: Kerollmops <clement@meilisearch.com>	2021-09-07 15:47:15 +00:00
Alexey Shekhirin	dbd91e7151	chore(ci): use smarter dependencies cache	2021-09-07 18:16:33 +03:00
bors[bot]	720becb5e8	Merge #341 341: Throw a query time error when a sort parameter is used but the sort ranking rule is missing r=Kerollmops a=Kerollmops This PR makes the engine throw an error for when the ranking rules don't contain the `sort` rule, the `sortable_fields` are correctly set but the user tries to use the `sort` query parameter. Doing so will have no effect on the returned documents so we preferred returning an error to help debug this. That's breaking on the MeiliSearch side as we added a new variant to the `UserError` enum. Co-authored-by: Kerollmops <clement@meilisearch.com>	2021-09-07 14:45:05 +00:00
Kerollmops	e2cefc9b4f	Move the sort ranking rule before the exactness ranking rule	2021-09-07 16:41:33 +02:00
bors[bot]	a0b3620b05	Merge #346 346: remove unused grenad default features r=Kerollmops a=MarinPostma Milli is not using any of grenad default features, and it's zstd feature creates conflict with meilisearch. This pr simply remove the unused features. Co-authored-by: mpostma <postma.marin@protonmail.com>	2021-09-07 14:30:20 +00:00
mpostma	cd043d4461	remove unused grenad default features	2021-09-07 16:21:46 +02:00
Kerollmops	5989528833	Add a test to make sure we throw the right error message	2021-09-07 11:02:00 +02:00
Kerollmops	fd3daa4423	Throw a query time error when a sort param is used but sort ranking rule is missing	2021-09-07 11:02:00 +02:00
Kerollmops	8dca36433c	Introduce the new SortRankingRuleMissing user error variant	2021-09-07 11:01:59 +02:00
bors[bot]	446ed17589	Merge #338 338: Fix string fields sorting r=Kerollmops a=shekhirin Resolves https://github.com/meilisearch/milli/issues/333 <details> <summary>curl checks</summary> ```console ➜ ~ curl -s 'localhost:7700/indexes/movies/search' -d '{"sort": ["title:asc"], "limit": 30}' \| jq -r '.hits \| map(.title)[]' #1 Cheerleader Camp #Horror #RealityHigh #Roxy #SquadGoals $5 a Day $9.99 '71 (2) (500) Days of Summer (Girl)Friend *batteries not included ...And God Created Woman ...And Justice for All ...E fuori nevica! .45 1 1 Mile To You 1 Night 10 10 Cloverfield Lane 10 giorni senza mamma 10 Items or Less 10 Rillington Place 10 Rules for Sleeping Around 10 Things I Hate About You 10 to Midnight 10 Years 10,000 BC 10,000 Saints ➜ ~ curl -s 'localhost:7700/indexes/movies/search' -d '{"sort": ["title:desc"], "limit": 30}' \| jq -r '.hits \| map(.title)[]' 크게 될 놈 왓칭 뷰티플 마인드 노무현과 바보들 ハニー Счастье – это… Часть 2 СОТКА Смотри мою любовь Позивний 'Бандерас' Лошо момиче Күлүк Хомус Куда течет море Каникулы президента Ακίνητο Ποτάμι Üç Harfliler: Beddua È nata una Star? Æon Flux Ága À propos de Nice À Nos Amours À l'aventure ¡Three Amigos! Zulu Dawn Zulu Zulu Zu: Warriors from the Magic Mountain Zu Warriors Zorro Zorba the Greek Zootopia ``` </details> Co-authored-by: Alexey Shekhirin <a.shekhirin@gmail.com>	2021-09-07 08:28:23 +00:00
Alexey Shekhirin	0be09555f1	test(search): asc/desc criteria for large datasets	2021-09-03 18:00:08 +03:00
Alexey Shekhirin	c2517e7d5f	fix(facet): string fields sorting	2021-09-03 11:58:26 +03:00
bors[bot]	5cbe879325	Merge #308 308: Implement a better parallel indexer r=Kerollmops a=ManyTheFish Rewrite the indexer: - enhance memory consumption control - optimize parallelism using rayon and crossbeam channel - factorize the different parts and make new DB implementation easier - optimize and fix prefix databases Co-authored-by: many <maxime@meilisearch.com>	2021-09-02 15:03:52 +00:00
many	741a4444a9	Remove log in chunk generator	2021-09-02 16:57:46 +02:00
many	7f7fafb857	Make document_chunk_size settable from update builder	2021-09-02 15:25:39 +02:00
many	db0c681bae	Fix Pr comments	2021-09-02 15:17:52 +02:00
bors[bot]	46f7df232a	Merge #337 337: Update version for the next release (v0.12.0) r=Kerollmops a=curquiza Breaking because of the new indexer that implies DB changes #308 Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2021-09-02 10:13:31 +00:00
Clémentine Urquizar	285849e3a6	Update version for the next release (v0.12.0)	2021-09-02 10:08:41 +02:00
bors[bot]	a589f6c60b	Merge #335 335: Get sortable_fields from index only if criteria present in query r=Kerollmops a=shekhirin Seems like we don't need to retrieve `sortable_fields` from the index if there's no any `sort_criteria` in the query. Small 🤏 optimization opportunity out there. Co-authored-by: Alexey Shekhirin <a.shekhirin@gmail.com>	2021-09-01 16:01:00 +00:00
bors[bot]	3e0a78acf3	Merge #329 329: Run all benchmarks once every friday r=irevoire a=irevoire All the benchmarks run every Friday on the `main` branch. To avoid having pending benchmarks everywhere, we execute one benchmark every 8 hours. Then the results are uploaded as if it was a normal user-run benchmark. This PR closes #314 and #321 Co-authored-by: Irevoire <tamo@meilisearch.com> Co-authored-by: Tamo <tamo@meilisearch.com>	2021-09-01 15:20:49 +00:00
many	4860fd4529	Ignore empty facet values	2021-09-01 16:48:40 +02:00
many	b3a22f31f6	Fix memory consuption in word pair proximity extractor	2021-09-01 16:48:40 +02:00
many	9452fabfb2	Optimize cbo roaring bitmaps merge	2021-09-01 16:48:40 +02:00
many	8f702828ca	Ignore errors comming from crossbeam channel senders	2021-09-01 16:48:40 +02:00
many	e09eec37bc	Handle distance addition with hard separators	2021-09-01 16:48:40 +02:00
many	fc7cc770d4	Add logging timers	2021-09-01 16:48:40 +02:00
many	a2f59a28f7	Remove unwrap sending errors in channel	2021-09-01 16:48:40 +02:00
many	5c962c03dd	Fix and optimize word_prefix_pair_proximity_docids database	2021-09-01 16:48:40 +02:00
many	2d1727697d	Take stop word in account	2021-09-01 16:48:40 +02:00
many	823da19745	Fix test and use progress callback	2021-09-01 16:48:39 +02:00
many	1d314328f0	Plug new indexer	2021-09-01 16:48:36 +02:00
many	3aaf1d62f3	Publish grenad CompressionType type in milli	2021-09-01 16:42:08 +02:00
Alexey Shekhirin	0e379558a1	fix(search): get sortable_fields only if criteria present	2021-08-31 21:35:41 +03:00
bors[bot]	d6bba0663a	Merge #334 334: Wrap long values into BStr for warn logs r=Kerollmops a=shekhirin Resolves https://github.com/meilisearch/milli/issues/263 Co-authored-by: Alexey Shekhirin <a.shekhirin@gmail.com>	2021-08-31 17:38:54 +00:00
Alexey Shekhirin	0b02eb456c	chore(update): wrap long values into BStr for warn logs	2021-08-31 20:28:16 +03:00
bors[bot]	df38794c7d	Merge #330 330: Introduce the reset_sortable_fields Settings method r=irevoire a=Kerollmops I forgot to add the `reset_sortable_fields` method on the `Settings` builder, it is no big deal as the library user (like MeiliSearch) can always call `set_sortable_fields` with an empty list of fields, it is equivalent. Co-authored-by: Kerollmops <clement@meilisearch.com> Co-authored-by: Tamo <tamo@meilisearch.com> Co-authored-by: bors[bot] <26634292+bors[bot]@users.noreply.github.com>	2021-08-30 23:26:54 +00:00
bors[bot]	6cdb6722d1	Merge #332 332: Sortable attributes in http-ui r=Kerollmops a=irevoire - Add a `reset_sortable_attribute` method - Add the `sortable_attributes` to http-ui - Fix some broken test in http-ui Co-authored-by: Tamo <tamo@meilisearch.com>	2021-08-30 15:31:05 +00:00
Tamo	d106eb5b90	add the sortable attributes to http-ui and fix the tests	2021-08-30 16:25:10 +02:00
Tamo	5e639bc0c1	postfix all action name with (cron)	2021-08-30 13:55:00 +02:00
Irevoire	49a6d2d5f1	run all benchmarks once every friday	2021-08-30 13:55:00 +02:00
Kerollmops	f230ae6fd5	Introduce the reset_sortable_fields Settings method	2021-08-25 17:44:16 +02:00
bors[bot]	c8930781eb	Merge #328 328: Remove `beta` compilation in CI r=Kerollmops a=shekhirin Resolves https://github.com/meilisearch/milli/issues/326 Co-authored-by: Alexey Shekhirin <a.shekhirin@gmail.com>	2021-08-25 08:45:18 +00:00
Alexey Shekhirin	01461af333	chore(ci): remove Rust beta from tests job	2021-08-24 22:18:13 +03:00
bors[bot]	c51bb6789c	Merge #325 325: Update milli version to v0.11.0 r=curquiza a=Kerollmops This PR also clean-up some dependencies in the Cargo.toml. Co-authored-by: Kerollmops <clement@meilisearch.com>	2021-08-24 16:18:49 +00:00
Kerollmops	af65485ba7	Reexport the grenad CompressionType from milli	2021-08-24 18:15:31 +02:00
Kerollmops	f2e1591826	Remove the unused tinytemplate dependency	2021-08-24 18:10:58 +02:00
Kerollmops	2f20257070	Update milli to the v0.11.0	2021-08-24 18:10:11 +02:00
bors[bot]	794c0f64a9	Merge #315 315: Rewrite the indexing benchmarks r=Kerollmops a=irevoire There was a panic on the benchmark and while I was trying to understand what was happening I decided to rewrite the way the benchmarks were working. Before we were creating a database with the good setting, and then for each benchmarks we were: 1. Deleting all documents in the database 2. Indexing a batch of documents Now for each iteration we recreate entirely a new database from scratch. Since deleting all the documents in a database may not be the same as starting with a fresh new database I prefer this solution. Co-authored-by: Irevoire <tamo@meilisearch.com>	2021-08-24 15:34:50 +00:00
bors[bot]	731e0e5321	Merge #320 320: Sort at query time r=Kerollmops a=Kerollmops Re-introduce the Sort at the query time (https://github.com/meilisearch/milli/issues/305) Co-authored-by: Clément Renault <renault.cle@gmail.com>	2021-08-24 14:19:43 +00:00
Clément Renault	89d0758713	Revert "Revert "Sort at query time""	2021-08-24 11:55:16 +02:00
bors[bot]	879d5e8799	Merge #319 319: Update version for the next release (v0.10.2) r=Kerollmops a=curquiza Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2021-08-23 10:03:23 +00:00
Clémentine Urquizar	88f6c18665	Update version for the next release (v0.10.2)	2021-08-23 11:33:30 +02:00
bors[bot]	aa1ce97748	Merge #317 317: Fix the facet string docids filterable deletion bug r=Kerollmops a=Kerollmops Fixes a bug where the deletion of documents was returning a decoding error. But only when the settings are set with filterable attributes. This bug was introduced in #254 in which we made the engine faster in returning the facet distribution. We changed the way we were storing the inverted index, we were no more storing only documents ids with the original values but also groups identified with integers, depending on the facet level we were using. This is similar to how facet numbers are already stored. ⚠️ As `@curquiza` already said, we must first revert #309 before merging this! Related to https://github.com/meilisearch/MeiliSearch/issues/1601. Co-authored-by: Clément Renault <clement@meilisearch.com>	2021-08-23 08:57:16 +00:00
Clément Renault	c084f7f731	Fix the facet string docids filterable deletion bug	2021-08-23 10:50:39 +02:00
bors[bot]	0d1f83ba4b	Merge #318 318: Revert "Sort at query time" r=Kerollmops a=curquiza Reverts meilisearch/milli#309 We revert this from `main` not because this leads to a bug, but because we don't want to release it now and we have to merge and release an hotfix on `main`. Cf: - https://github.com/meilisearch/milli/issues/316 - https://github.com/meilisearch/milli/pull/317 Once the v0.21.0 is released, we should merge again this awesome addition 👌 Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2021-08-21 08:25:17 +00:00
Clémentine Urquizar	922f9fd4d5	Revert "Sort at query time"	2021-08-20 18:09:17 +02:00
Irevoire	4b99d8cb91	rewrite the indexing benchmarks	2021-08-19 15:02:43 +02:00
bors[bot]	41fc0dcb62	Merge #309 309: Sort at query time r=Kerollmops a=Kerollmops This PR: - Makes the `Asc/Desc` criteria work with strings too, it first returns documents ordered by numbers then by strings, and finally the documents that can't be ordered. Note that it is lexicographically ordered and not ordered by character, which means that it doesn't know about wide and short characters i.e. `a`, `丹`, `▲`. - Changes the syntax for the `Asc/Desc` criterion by now using a colon to separate the name and the order i.e. `title:asc`, `price:desc`. - Add the `Sort` criterion at the third position in the ranking rules by default. - Add the `sort_criteria` method to the `Search` builder struct to let the users define the `Asc/Desc` sortable attributes they want to use at query time. Note that we need to check that the fields are registered in the sortable attributes before performing the search. - Introduce a new `InvalidSortableAttribute` user error that is raised when the sort criteria declared at query time are not part of the sortable attributes. - `@ManyTheFish` introduced integration tests for the dynamic Sort criterion. Fixes #305. Co-authored-by: Kerollmops <clement@meilisearch.com> Co-authored-by: many <maxime@meilisearch.com>	2021-08-18 16:55:32 +00:00
many	d1df0d20f9	Add integration test of SortBy criterion	2021-08-18 16:21:51 +02:00
Kerollmops	1b7f6ea1e7	Return a new error when the sort criteria is not sortable	2021-08-18 15:04:07 +02:00
Kerollmops	71602e0f1b	Add the sortable fields into the settings and in the index	2021-08-18 15:04:07 +02:00
Kerollmops	407f53872a	Add a sort_criteria method to the Search builder struct	2021-08-18 15:04:07 +02:00
Kerollmops	687cd2e205	Introduce the new Sort criterion and AscDesc enum	2021-08-18 15:04:07 +02:00
bors[bot]	198c416bd8	Merge #312 312: Update milli version to v0.10.1 r=Kerollmops a=curquiza Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2021-08-18 12:08:04 +00:00
Clémentine Urquizar	6cb9c3b81f	Update milli version to v0.10.1	2021-08-18 13:46:27 +02:00
bors[bot]	2a67308e29	Merge #311 311: Update tokenizer version to v0.2.5 r=Kerollmops a=curquiza Fixes panic when indexing data containing [control characters](https://en.wikipedia.org/wiki/Control_character) but continue accepting whitespace, obviously. Related to https://github.com/meilisearch/MeiliSearch/issues/1590 Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2021-08-18 11:41:52 +00:00
Clémentine Urquizar	42cf847a63	Update tokenizer version to v0.2.5	2021-08-18 13:37:41 +02:00
bors[bot]	c4275f0d27	Merge #310 310: Modify the README file r=Kerollmops a=Kerollmops Co-authored-by: Kerollmops <clement@meilisearch.com>	2021-08-17 15:20:43 +00:00
Kerollmops	ecf8abc518	Modify the README file	2021-08-17 17:18:58 +02:00
Kerollmops	5b88df508e	Use the new Asc/Desc syntax everywhere	2021-08-17 14:15:22 +02:00
Kerollmops	fcedff95e8	Change the Asc/Desc criterion syntax to use a colon (:)	2021-08-17 14:03:21 +02:00
Kerollmops	e9ada44509	AscDesc criterion returns documents ordered by numbers then by strings	2021-08-17 13:21:31 +02:00
Kerollmops	110bf6b778	Make the FacetStringIter work in both, ascending and descending orders	2021-08-17 11:18:40 +02:00
Kerollmops	22ebd2658f	Introduce the EitherString/RevRange private aliases	2021-08-17 10:47:15 +02:00
Kerollmops	7a5889bc5a	Introduce the highest_reverse_iter private method	2021-08-17 10:45:26 +02:00
Kerollmops	ad0d311f8a	Introduce the FacetStringLevelZeroRevRange struct	2021-08-17 10:44:43 +02:00
Kerollmops	6214c38da9	Introduce the FacetStringGroupRevRange struct	2021-08-17 10:44:27 +02:00
Kerollmops	1c604de158	Introduce the highest_iter private method on the FacetStringIter struct	2021-08-17 10:41:11 +02:00
Kerollmops	64df159057	Introduce the new_reducing constructor on the FacetStringIter struct	2021-08-17 10:35:06 +02:00
Kerollmops	01a4052828	Move the FacetStringIter creation logic into a private new method	2021-08-17 10:29:43 +02:00
bors[bot]	51581d14f8	Merge #307 307: Update version for the next release (v0.10.0) r=Kerollmops a=curquiza Replaces https://github.com/meilisearch/milli/pull/304 Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2021-08-16 10:33:53 +00:00
Clémentine Urquizar	fcc520e49a	Update version for the next release (v0.10.0)	2021-08-16 12:00:28 +02:00
bors[bot]	1541bce952	Merge #303 303: Remove max values by facet limit for facet distribution r=Kerollmops a=ManyTheFish Co-authored-by: many <maxime@meilisearch.com>	2021-08-16 09:58:53 +00:00
many	7dbefae1e3	Make facet string iterator non reducing	2021-08-12 17:23:39 +02:00
many	8fdf860c17	Remove max values by facet limit for facet distribution	2021-08-12 11:29:20 +02:00
bors[bot]	2102e0da6b	Merge #302 302: Update milli to v0.9.0 r=curquiza a=curquiza Updating the minor and not patch since #300 seems to be breaking: it involves a re-indexation to get the fix, so it involves an additional step from the users, not only downloading the latest version. Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2021-08-05 08:38:15 +00:00
bors[bot]	89b9b61840	Merge #300 300: Fix prefix level position docids database r=curquiza a=ManyTheFish The prefix search was inverted when we generated the DB. Instead of searching if word had a prefix in prefix fst, we were searching if the word was a prefix of a prefix contained in the prefix fst. The indexer, now, iterate over prefix contained in the fst and search them by prefix in the word-level-position-docids database, aggregating matches in a sorter. Fix #299 Co-authored-by: many <maxime@meilisearch.com>	2021-08-04 16:52:09 +00:00
Clémentine Urquizar	7f26c75610	Update milli to v0.9.0	2021-08-04 16:04:55 +02:00
many	cdeb07f0fd	Fix prefix level position docids database The prefix search was inverted when we generated the DB. Instead of searching if word had a prefix in prefix fst, we were searching if the word was a prefix of a prefix contained in the prefix fst. The indexer, now, iterate over prefix contained in the fst and search them by prefix in the word-level-position-docids database, aggregating matches in a sorter. Fix #299	2021-08-04 14:11:49 +02:00
bors[bot]	cb45a10bcd	Merge #298 298: Rename the search benchmarks r=Kerollmops a=irevoire And fix a bug. As always, I was not closing the env. Co-authored-by: Tamo <tamo@meilisearch.com>	2021-07-29 15:33:15 +00:00
Tamo	7eb2d71009	fix the benchmarks	2021-07-29 16:27:05 +02:00
Tamo	976dc1f4bc	prefix the search benchmarks with 'search'	2021-07-29 16:27:05 +02:00
bors[bot]	1290edd58a	Merge #297 297: Bump milli to v0.8.1 r=curquiza a=Kerollmops Co-authored-by: Kerollmops <clement@meilisearch.com>	2021-07-29 14:19:41 +00:00
Kerollmops	341c244965	Bump milli to v0.8.1	2021-07-29 15:56:36 +02:00
bors[bot]	d962e46ed1	Merge #296 296: Fix invalid faceted documents ids buffer size r=Kerollmops a=Kerollmops Fix a bug found by `@irevoire` when benchmarking the search. Co-authored-by: Kerollmops <clement@meilisearch.com>	2021-07-29 13:52:34 +00:00
Kerollmops	90514e03d1	Fix invalid faceted documents ids buffer size	2021-07-29 15:49:23 +02:00
bors[bot]	200e98c211	Merge #293 293: Make sure that the relevancy is not impacted by other settings r=Kerollmops a=Kerollmops Fix https://github.com/meilisearch/meilisearch/issues/1505. fix https://github.com/meilisearch/MeiliSearch/issues/1529 Co-authored-by: Kerollmops <clement@meilisearch.com>	2021-07-27 16:04:52 +00:00
bors[bot]	bc845324df	Merge #295 295: Update version for the next release (v0.8.0) r=Kerollmops a=curquiza Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2021-07-27 14:42:10 +00:00
Clémentine Urquizar	6a141694da	Update version for the next release (v0.8.0)	2021-07-27 16:38:42 +02:00
Kerollmops	dc2b63abdf	Introduce an empty FilterCondition variant to support unknown fields	2021-07-27 16:34:04 +02:00
bors[bot]	4ab7ca0e83	Merge #288 288: Stop tracking the Cargo.lock and add cache + windows to the CI r=curquiza a=irevoire We reuse the same `~/.cargo` and `./target` directory between each run on the same OS and rust toolchain. The `key` to decide if we can use the cache or not is: `$OS_NAME-$RUST_TOOLCHAIN-$HASH(Cargo.toml)` We also removed the `Cargo.lock` from this repository. Indeed, milli is a library and [should not track the `Cargo.lock`](https://doc.rust-lang.org/cargo/guide/cargo-toml-vs-cargo-lock.html) And finally, we enabled the tests on `windows-latest`. Since `lmdb` has been updated, this is now possible. Co-authored-by: Tamo <tamo@meilisearch.com>	2021-07-26 14:22:19 +00:00
Tamo	0038b3848a	add a simple github cache	2021-07-26 15:31:26 +02:00
Tamo	88646a63a1	update bors	2021-07-26 15:31:00 +02:00
Kerollmops	b12738cfe9	Use the right DB prefixes to store the faceted fields	2021-07-22 19:18:22 +02:00
Kerollmops	7aa6cc9b04	Do not insert fields in the map when changing the settings	2021-07-22 18:40:12 +02:00
bors[bot]	ee3a49cfba	Merge #291 291: Fix a bug about zero bytes in the inputs r=irevoire a=Kerollmops Ok, good news, after a little session of debugging with `@irevoire` we found out that the bug seems to be related to zeroes in the input update. The engine wasn't designed to accept those. The chosen solution is to update the tokenizer to remove those zeroes. We are waiting on https://github.com/meilisearch/tokenizer/pull/52 to be merged and a new version to be released. It is not an undefined behavior, I repeat: it is a "normal" bug 🎉 👏 ---- This PR tries to fix a bug where we use LMDB in the wrong way, leading to panic due to an undefined behavior on the Rust side. I thought [we fixed it in a previous PR](https://github.com/meilisearch/milli/pull/264) but we found out that _a similar_ bug was still present. `@bb` found a way to trigger this bug and helped us find the origin of it. As I don't have a minimal reproducible example of this bug I bet on the unsafe `put_current` calls when we index new documents as the bug was trigger after a big indexation on a clean database, thus not triggering a deletion update. I only replaced the unsafe `put_current` with two safe calls to `get`/`put`. I hope it helps and fixes the bug, only `@bb` can help us check that. I am not even sure how I can create a custom Docker image and expose it for testing purposes. <details> <summary>The backtrace leading us to a panic in grenad.</summary> ``` meilisearch_1 \| thread 'tokio-runtime-worker' panicked at 'assertion failed: key > &last_key', /root/.cargo/git/checkouts/grenad-e2cb77f65d31bb02/3adcb26/src/block_builder.rs:38:17 meilisearch_1 \| stack backtrace: meilisearch_1 \| 0: rust_begin_unwind meilisearch_1 \| at ./rustc/53cb7b09b00cbea8754ffb78e7e3cb521cb8af4b/library/std/src/panicking.rs:493:5 meilisearch_1 \| 1: core::panicking::panic_fmt meilisearch_1 \| at ./rustc/53cb7b09b00cbea8754ffb78e7e3cb521cb8af4b/library/core/src/panicking.rs:92:14 meilisearch_1 \| 2: core::panicking::panic meilisearch_1 \| at ./rustc/53cb7b09b00cbea8754ffb78e7e3cb521cb8af4b/library/core/src/panicking.rs:50:5 meilisearch_1 \| 3: grenad::block_builder::BlockBuilder::insert meilisearch_1 \| at ./root/.cargo/git/checkouts/grenad-e2cb77f65d31bb02/3adcb26/src/block_builder.rs:38:17 meilisearch_1 \| 4: grenad::writer::Writer<W>::insert meilisearch_1 \| at ./root/.cargo/git/checkouts/grenad-e2cb77f65d31bb02/3adcb26/src/writer.rs:92:12 meilisearch_1 \| 5: milli::update::words_level_positions::write_level_entry meilisearch_1 \| at ./root/.cargo/git/checkouts/milli-00376cd5db949a15/007fec2/milli/src/update/words_level_positions.rs:262:5 meilisearch_1 \| 6: milli::update::words_level_positions::compute_positions_levels meilisearch_1 \| at ./root/.cargo/git/checkouts/milli-00376cd5db949a15/007fec2/milli/src/update/words_level_positions.rs:211:13 meilisearch_1 \| 7: milli::update::words_level_positions::WordsLevelPositions::execute meilisearch_1 \| at ./root/.cargo/git/checkouts/milli-00376cd5db949a15/007fec2/milli/src/update/words_level_positions.rs:65:23 meilisearch_1 \| 8: milli::update::index_documents::IndexDocuments::execute_raw meilisearch_1 \| at ./root/.cargo/git/checkouts/milli-00376cd5db949a15/007fec2/milli/src/update/index_documents/mod.rs:831:9 meilisearch_1 \| 9: milli::update::index_documents::IndexDocuments::execute meilisearch_1 \| at ./root/.cargo/git/checkouts/milli-00376cd5db949a15/007fec2/milli/src/update/index_documents/mod.rs:372:9 meilisearch_1 \| 10: meilisearch_http::index::updates::<impl meilisearch_http::index::Index>::update_documents_txn meilisearch_1 \| at ./meilisearch/meilisearch-http/src/index/updates.rs:225:30 meilisearch_1 \| 11: meilisearch_http::index::updates::<impl meilisearch_http::index::Index>::update_documents meilisearch_1 \| at ./meilisearch/meilisearch-http/src/index/updates.rs:183:22 meilisearch_1 \| 12: meilisearch_http::index::update_handler::UpdateHandler::handle_update meilisearch_1 \| at ./meilisearch/meilisearch-http/src/index/update_handler.rs:75:18 meilisearch_1 \| 13: meilisearch_http::index_controller::index_actor::actor::IndexActor<S>::handle_update::{{closure}}::{{closure}} meilisearch_1 \| at ./meilisearch/meilisearch-http/src/index_controller/index_actor/actor.rs:174:35 meilisearch_1 \| 14: <tokio::runtime::blocking::task::BlockingTask<T> as core::future::future::Future>::poll meilisearch_1 \| at ./root/.cargo/registry/src/github.com-1ecc6299db9ec823/tokio-1.7.1/src/runtime/blocking/task.rs:42:21 meilisearch_1 \| 15: tokio::runtime::task::core::CoreStage<T>::poll::{{closure}} meilisearch_1 \| at ./root/.cargo/registry/src/github.com-1ecc6299db9ec823/tokio-1.7.1/src/runtime/task/core.rs:243:17 meilisearch_1 \| 16: tokio::loom::std::unsafe_cell::UnsafeCell<T>::with_mut meilisearch_1 \| at ./root/.cargo/registry/src/github.com-1ecc6299db9ec823/tokio-1.7.1/src/loom/std/unsafe_cell.rs:14:9 meilisearch_1 \| 17: tokio::runtime::task::core::CoreStage<T>::poll meilisearch_1 \| at ./root/.cargo/registry/src/github.com-1ecc6299db9ec823/tokio-1.7.1/src/runtime/task/core.rs:233:13 meilisearch_1 \| 18: tokio::runtime::task::harness::poll_future::{{closure}} meilisearch_1 \| at ./root/.cargo/registry/src/github.com-1ecc6299db9ec823/tokio-1.7.1/src/runtime/task/harness.rs:427:23 meilisearch_1 \| 19: <std::panic::AssertUnwindSafe<F> as core::ops::function::FnOnce<()>>::call_once meilisearch_1 \| at ./rustc/53cb7b09b00cbea8754ffb78e7e3cb521cb8af4b/library/std/src/panic.rs:344:9 meilisearch_1 \| 20: std::panicking::try::do_call meilisearch_1 \| at ./rustc/53cb7b09b00cbea8754ffb78e7e3cb521cb8af4b/library/std/src/panicking.rs:379:40 meilisearch_1 \| 21: std::panicking::try meilisearch_1 \| at ./rustc/53cb7b09b00cbea8754ffb78e7e3cb521cb8af4b/library/std/src/panicking.rs:343:19 meilisearch_1 \| 22: std::panic::catch_unwind meilisearch_1 \| at ./rustc/53cb7b09b00cbea8754ffb78e7e3cb521cb8af4b/library/std/src/panic.rs:431:14 meilisearch_1 \| 23: tokio::runtime::task::harness::poll_future meilisearch_1 \| at ./root/.cargo/registry/src/github.com-1ecc6299db9ec823/tokio-1.7.1/src/runtime/task/harness.rs:414:19 meilisearch_1 \| 24: tokio::runtime::task::harness::Harness<T,S>::poll_inner meilisearch_1 \| at ./root/.cargo/registry/src/github.com-1ecc6299db9ec823/tokio-1.7.1/src/runtime/task/harness.rs:89:9 meilisearch_1 \| 25: tokio::runtime::task::harness::Harness<T,S>::poll meilisearch_1 \| at ./root/.cargo/registry/src/github.com-1ecc6299db9ec823/tokio-1.7.1/src/runtime/task/harness.rs:59:15 meilisearch_1 \| 26: tokio::runtime::task::raw::RawTask::poll meilisearch_1 \| at ./root/.cargo/registry/src/github.com-1ecc6299db9ec823/tokio-1.7.1/src/runtime/task/raw.rs:66:18 meilisearch_1 \| 27: tokio::runtime::task::Notified<S>::run meilisearch_1 \| at ./root/.cargo/registry/src/github.com-1ecc6299db9ec823/tokio-1.7.1/src/runtime/task/mod.rs:171:9 meilisearch_1 \| 28: tokio::runtime::blocking::pool::Inner::run meilisearch_1 \| at ./root/.cargo/registry/src/github.com-1ecc6299db9ec823/tokio-1.7.1/src/runtime/blocking/pool.rs:265:17 meilisearch_1 \| 29: tokio::runtime::blocking::pool::Spawner::spawn_thread::{{closure}} meilisearch_1 \| at ./root/.cargo/registry/src/github.com-1ecc6299db9ec823/tokio-1.7.1/src/runtime/blocking/pool.rs:245:17 meilisearch_1 \| note: Some details are omitted, run with `RUST_BACKTRACE=full` for a verbose backtrace. ``` </details> Co-authored-by: Kerollmops <clement@meilisearch.com>	2021-07-22 16:14:35 +00:00
Kerollmops	0353fbb5df	Bump the tokenizer version to v0.2.4	2021-07-22 17:14:45 +02:00
Kerollmops	92c0a2cdc1	Add a test that triggers a panic when indexing zeroes	2021-07-22 17:14:44 +02:00
Kerollmops	aa02a7fdd8	Add a test to check that we indeed impact the relevancy	2021-07-22 17:04:38 +02:00
bors[bot]	77de82aaa4	Merge #254 254: Improve the facet string distribution speed r=Kerollmops a=Kerollmops This pull request creates a data structure similar to the one we use for the faceted numbers, a tetratomic decision tree but this time for the facet strings. This PR also changes the facet distribution behavior by returning one of the original facet values, fixes #260. This data structure defines bucket-like structures where documents ids are stored under their facet value and helps the search decide if it wants to move to a lower level under a given bucket or not, depending on if the current bucket contains interesting documents or not. The whole format, algorithm, and previous attempts are explained in the [`facet_string.rs` file](`ec1cfdd42b/milli/src/search/facet/facet_string.rs`). Note that this data structure could be used to sort by string lexicographically, that hypothetically possible. We need more testing, in terms of performance and quality, as we will sort on lowercased versions of the facet values. - [x] Implement a faster and more precise way to fetch the facet distribution. - [x] Store and return the original facet string value. We currently return the lowercased version. Co-authored-by: Kerollmops <clement@meilisearch.com> Co-authored-by: Clément Renault <clement@meilisearch.com>	2021-07-21 15:34:40 +00:00
Clément Renault	0227254a65	Return the original string values for the inverted facet index database	2021-07-21 16:59:39 +02:00
Kerollmops	03a01166ba	Display the original facet string value from the linear facet database	2021-07-21 16:59:39 +02:00
Clément Renault	d23c250ad5	Fix a bound error in the facet string range construction	2021-07-21 16:59:39 +02:00
Clément Renault	081278dfd6	Use the facet string levels when computing the facet distribution	2021-07-21 16:59:39 +02:00
Clément Renault	5676b204dd	Fix the facet string levels codecs	2021-07-21 16:59:38 +02:00
Kerollmops	8c86348119	Indexing the facet strings levels	2021-07-21 16:59:38 +02:00
Kerollmops	a7ae552ba7	Fix the FacetStringLevelZeroRange range when unbounded	2021-07-21 16:59:38 +02:00
Kerollmops	757b2b502a	Remove the FacetValueStringCodec	2021-07-21 16:59:38 +02:00
Kerollmops	adfd4da24c	Introduce the FacetStringIter iterator	2021-07-21 16:59:38 +02:00
Kerollmops	a79661c6dc	Introduce a lot of facet string helper iterators	2021-07-21 16:59:38 +02:00
Kerollmops	851f979039	Describe the way we want to group the facet strings	2021-07-21 16:59:38 +02:00
Kerollmops	f858f64b1f	Move the facet number iterators into their own module	2021-07-21 16:59:37 +02:00
Kerollmops	9f8095c069	Make sure that we don't keep a reference on the LMDB key when using put_current	2021-07-21 10:35:35 +02:00
bors[bot]	fa44e95c91	Merge #290 290: Add a $HOME to the CI r=Kerollmops a=irevoire This should fix this issue: https://github.com/meilisearch/milli/runs/3104228432?check_suite_focus=true I think a real fix would be to fix the configuration of our github runner but I don't know how to do it. @curquiza could probably help us on that once she's back from vacation 😄 Co-authored-by: Tamo <tamo@meilisearch.com>	2021-07-20 07:32:46 +00:00
Tamo	0ab541627b	add a $HOME var to the ci	2021-07-19 14:33:49 +02:00
bors[bot]	16698f714b	Merge #287 287: Add benchmarks for indexing r=Kerollmops a=irevoire closes #274 I don't really know how much time this will take on our bench machine. I'm afraid the wiki dataset will take a really long time to bench (it takes 1h30 on my computer). If you are ok with it, I would like to merge this first PR since it introduces a first set of benchmarks and see how much time it takes in reality on our setup. Co-authored-by: Tamo <tamo@meilisearch.com>	2021-07-07 15:41:15 +00:00
Tamo	931021fe57	add benchmarks for indexing	2021-07-07 13:09:05 +02:00
bors[bot]	4c9531bdf3	Merge #285 285: Support documents with at most 65536 fields r=Kerollmops a=Kerollmops Fixes #248. In this PR I updated the `obkv` crate, it now supports arbitrary key length and therefore I was able to use an `u16` to represent the fields instead of a single byte. It was impressively easy to update the whole codebase 🍡 🍔 Co-authored-by: Kerollmops <clement@meilisearch.com>	2021-07-06 16:44:51 +00:00
Kerollmops	0a78107525	Fix the infos crate to make it read u16 field ids	2021-07-06 11:58:03 +02:00
Kerollmops	a9553af635	Add a test to check that we can index more that 256 fields	2021-07-06 11:58:03 +02:00
Kerollmops	838ed1cd32	Use an u16 field id instead of one byte	2021-07-06 11:58:03 +02:00
bors[bot]	cc54c41e30	Merge #283 283: Use the AlwaysFreePages flag when opening an index r=irevoire a=Kerollmops We introduced a new flag in our fork of LMDB, this `AlwaysFreePages` flag forces LMDB to always free the single pages it uses before writing to the disk instead of keeping them in a linked list. Declaring this flag reduces the memory print (leak) we have on memory after indexing a lot of documents. Fixes #279. Co-authored-by: Kerollmops <clement@meilisearch.com>	2021-07-05 16:59:16 +00:00
bors[bot]	63db43cc7a	Merge #284 284: [http-ui] Introduce the route `die` r=Kerollmops a=irevoire This route just `exit` the process. This can come in handy when you run `http-ui` inside of another process (a profiler for example), and you don't want to kill everything Co-authored-by: Tamo <tamo@meilisearch.com> Co-authored-by: Irevoire <tamo@meilisearch.com>	2021-07-05 15:47:53 +00:00
Irevoire	4562b278a8	remove a warning and add a log Co-authored-by: Clément Renault <clement@meilisearch.com>	2021-07-05 17:46:02 +02:00
Tamo	a57e522a67	introduce a die route let the program exit itself alone	2021-07-05 17:38:10 +02:00
Kerollmops	91c5d0c042	Use the AlwaysFreePages flag when opening an index	2021-07-05 16:36:13 +02:00
bors[bot]	007fec21fc	Merge #281 281: Bump to v0.7.2 r=ManyTheFish a=Kerollmops Co-authored-by: Kerollmops <clement@meilisearch.com>	2021-07-05 09:00:26 +00:00
Kerollmops	a6b4069172	Bump to v0.7.2	2021-07-05 10:54:53 +02:00
bors[bot]	d7bc6a6999	Merge #280 280: Fix matching lenghth in matching_words r=Kerollmops a=ManyTheFish related to https://github.com/meilisearch/MeiliSearch/issues/1441 Co-authored-by: many <maxime@meilisearch.com>	2021-07-01 18:50:46 +00:00
many	9f62149b94	Fix matching lenghth in matching_words	2021-07-01 19:03:28 +02:00
bors[bot]	f25f454bd4	Merge #275 275: Fix the benchmarks dependencies r=Kerollmops a=irevoire Import exactly the same dependency as milli instead of a wildcard that can do anything Co-authored-by: Tamo <tamo@meilisearch.com> Co-authored-by: Irevoire <irevoire@protonmail.ch>	2021-07-01 11:07:01 +00:00
bors[bot]	885f243afc	Merge #276 276: Fix the fmt of the auto-generated file r=Kerollmops a=irevoire The file generated by the `build.rs` file of the benchmark was badly formatted and that was causing an issue with the git pre-commit hook I wrote [earlier](https://github.com/meilisearch/milli/blob/main/script/pre-commit) Co-authored-by: Tamo <tamo@meilisearch.com>	2021-07-01 10:24:36 +00:00
Irevoire	ec87bf3dd5	Update benchmarks/Cargo.toml Co-authored-by: Clément Renault <renault.cle@gmail.com>	2021-07-01 11:45:05 +02:00
Tamo	ef965aa3f3	fix the fmt of the auto-generated file	2021-07-01 11:43:09 +02:00
Tamo	fc09d77e89	fix the benchmarks dependcies	2021-07-01 11:38:30 +02:00
bors[bot]	056180e6c8	Merge #273 273: Update tokenizer version to v0.2.3 r=Kerollmops a=curquiza Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2021-07-01 09:02:16 +00:00
Clémentine Urquizar	3c149d8a43	Update tokenizer version to v0.2.3	2021-06-30 18:41:35 +02:00
bors[bot]	b4dcdbf00d	Merge #269 #271 269: Fix bug when inserting previously deleted documents r=Kerollmops a=Kerollmops This PR fixes #268. The issue was in the `ExternalDocumentsIds` implementation in the specific case that an external document id was in the soft map marked as deleted. The bug was due to a wrong assumption on my side about how the FST unions were returning the `IndexedValue`s, I thought the values returned in an array were in the same order as the FSTs given to the `OpBuilder` but in fact, [the `IndexedValue`'s `index` field was here to indicate from which FST the values were coming from](https://docs.rs/fst/0.4.7/fst/map/struct.IndexedValue.html). 271: Remove the roaring operation functions warnings r=Kerollmops a=Kerollmops In this PR we are just replacing the usages of the roaring operations function by the new operators. This removes a lot of warnings. Co-authored-by: Kerollmops <clement@meilisearch.com>	2021-06-30 12:34:55 +00:00
Kerollmops	32b7bd366f	Remove the roaring operation functions warnings	2021-06-30 14:12:56 +02:00
bors[bot]	00e2845f0f	Merge #270 270: Update milli version to v0.7.1 r=Kerollmops a=curquiza Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2021-06-30 12:12:24 +00:00
Kerollmops	c92ef54466	Add a test for when we insert a previously deleted document	2021-06-30 14:00:01 +02:00
Kerollmops	28782ff99d	Fix ExternalDocumentsIds struct when inserting previously deleted ids	2021-06-30 14:00:01 +02:00
Clémentine Urquizar	b489515f4d	Update milli version to v0.7.1	2021-06-30 13:52:46 +02:00
Kerollmops	54889813ce	Implement some debug functions on the ExternalDocumentsIds struct	2021-06-30 11:29:41 +02:00
Kerollmops	4bce66d5ff	Make the Index::delete_* method private	2021-06-30 10:07:31 +02:00
bors[bot]	66e6ea56b8	Merge #267 267: Highlighting r=Kerollmops a=irevoire closes #262 I basically rewrote a part of the damerau-levenshtein function we were using for the highlighting to accept at most two errors from the user and stop on the third mistake. Also, now it supports utf-8, so it should fix our issue. Co-authored-by: Tamo <tamo@meilisearch.com> Co-authored-by: Irevoire <irevoire@protonmail.ch>	2021-06-30 05:43:50 +00:00
Irevoire	6044b80362	Update milli/src/search/matching_words.rs Co-authored-by: Clément Renault <renault.cle@gmail.com>	2021-06-30 00:35:26 +02:00
Tamo	be75e738b1	add more tests	2021-06-29 16:24:58 +02:00
Tamo	56fceb1928	re-implement the Damerau-Levenshtein used for the highlighting	2021-06-29 15:36:03 +02:00
bors[bot]	9dbc8b2dd0	Merge #266 266: Bump LMDB to the latest version (v0.9.70) r=Kerollmops a=Kerollmops By bumping to a new version of heed (from git, v0.12.0 unpublished yet), this PR fixes Windows disk reservation problems. This new version of heed changes the `del/put_current`, and `append` iterator methods signature by declaring them unsafe. This PR also bumps milli itself into v0.7.0 as it is breaking due to the heed/LMDB bump. This PR must be merged after #264. Co-authored-by: Clément Renault <clement@meilisearch.com>	2021-06-28 17:11:41 +00:00
Clément Renault	80c6aaf1fd	Bump milli to 0.7.0	2021-06-28 18:31:56 +02:00
Clément Renault	bdc5599b73	Bump heed to use the git repo with v0.12.0	2021-06-28 18:26:20 +02:00
Clément Renault	73384aec21	Merge pull request #264 from meilisearch/fix-heed-undefined-behavior Fix the invalid heed usage	2021-06-28 18:23:49 +02:00
Clément Renault	0013236e5d	Fix the LMDB and heed invalid interactions. It is undefined behavior to keep a reference to the database while modifying it, we were keeping references in the database and also feeding the heed put_current methods with keys referenced inside the database itself. https://github.com/Kerollmops/heed/pull/108	2021-06-28 16:19:02 +02:00
Kerollmops	9e5f9a8a10	Add a test for the words level positions generation bug	2021-06-28 16:08:31 +02:00
bors[bot]	c38b0b883d	Merge #257 257: Fix unconditional facet indexing r=Kerollmops a=Kerollmops We were indexing every searchable field as filterable, this was a mistake. Co-authored-by: Kerollmops <clement@meilisearch.com>	2021-06-23 15:32:46 +00:00
Kerollmops	98285b4b18	Bump milli to 0.6.0	2021-06-23 17:30:26 +02:00
Kerollmops	4fc8f06791	Rename faceted_fields into filterable_fields	2021-06-23 17:26:54 +02:00
Kerollmops	c31cadb54f	Do not consider the searchable field as filterable	2021-06-23 17:26:54 +02:00
bors[bot]	41c4a5b60d	Merge #246 246: Improve the ci r=Kerollmops a=irevoire Rewrite the CI entirely: - run the ci on Linux, macOS and Windows. - run the ci on rust stable, beta and nightly - add rustfmt to the CI. - split the CI into multiple tasks, this way, the ci should be faster to fail Co-authored-by: Tamo <tamo@meilisearch.com> Co-authored-by: Irevoire <tamo@meilisearch.com>	2021-06-23 12:52:39 +00:00
Irevoire	faa3cd3b71	Update bors.toml Don't check nightly and beta channel Co-authored-by: Clément Renault <clement@meilisearch.com>	2021-06-23 14:30:33 +02:00
bors[bot]	2ab24c4f49	Merge #256 256: Update version for the next release (v0.5.1) r=Kerollmops a=curquiza Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2021-06-23 12:29:57 +00:00
Clémentine Urquizar	9885fb4159	Update version for the next release (v0.5.1)	2021-06-23 14:05:20 +02:00
bors[bot]	66f55e3e6a	Merge #255 255: Fix facet distribution error r=Kerollmops a=Kerollmops This PR fixes two invalid behaviors and fixes #253: - We were ignoring the list of fields for which the user wanted a facet distribution. - We were not raising any error for when a non-filterable field was requested a facet distribution. ~For the latter behavior I need the help of @curquiza to help me choose the right error type.~ Co-authored-by: Kerollmops <clement@meilisearch.com>	2021-06-23 12:03:05 +00:00
Kerollmops	a6218a20ae	Introduce a new InvalidFacetsDistribution user error	2021-06-23 13:56:19 +02:00
Kerollmops	2364777838	Return an error for when a field distribution cannot be done	2021-06-23 11:50:49 +02:00
Kerollmops	aeaac743ff	Replace an if let some by a match	2021-06-23 11:33:30 +02:00
Tamo	5099192c44	update bors.toml	2021-06-23 10:22:40 +02:00
Tamo	d8695da1d1	improve the ci	2021-06-23 10:22:40 +02:00
bors[bot]	28197b2435	Merge #252 252: Run the formatter on the whole project a second time r=Kerollmops a=irevoire Co-authored-by: Tamo <tamo@meilisearch.com>	2021-06-22 13:56:09 +00:00
Tamo	8d2a0b43ff	run the formatter on the whole project a second time	2021-06-22 15:36:22 +02:00
bors[bot]	634201244c	Merge #250 #251 250: Add the limit field to http-ui r=Kerollmops a=irevoire 251: Fix the limit r=Kerollmops a=irevoire There was no check on the limit and thus if a user specified a very large number this line could cause a panic. Co-authored-by: Tamo <tamo@meilisearch.com>	2021-06-22 13:00:52 +00:00
Tamo	3d90b03d7b	fix the limit There was no check on the limit and thus, if a user especified a very large number this line could causes a panic	2021-06-22 14:52:13 +02:00
Tamo	81643e6d70	add the limit field to http-ui	2021-06-22 14:47:23 +02:00
bors[bot]	5aea8dd75b	Merge #249 249: Enable the jemallocator dependencies only when we are running on linux r=Kerollmops a=irevoire Co-authored-by: Tamo <tamo@meilisearch.com>	2021-06-22 12:32:44 +00:00
Tamo	77eb37934f	add jemalloc to http-ui and the benchmarks	2021-06-22 14:17:56 +02:00
bors[bot]	5b6adc6d96	Merge #245 245: Warn for when a key is too large for LMDB r=Kerollmops a=Kerollmops Closes #191, and resolves #140. Co-authored-by: Kerollmops <clement@meilisearch.com>	2021-06-22 12:10:52 +00:00
Tamo	d53df8a002	enable the jemallocator dependencies only when we are running on linux	2021-06-22 14:04:16 +02:00
bors[bot]	ca9fa329d1	Merge #247 247: Return a `MissingDocumentId` error when a document doesn't have one r=Kerollmops a=Kerollmops We were wrongly returning a `MissingPrimaryKey` instead of a `MissingDocumentId` error for when a document was missing a document id. We also improved the error message for when a document id is invalid (wrong type or wrong format). Co-authored-by: Kerollmops <clement@meilisearch.com>	2021-06-22 10:07:54 +00:00
Kerollmops	51dbb2e06d	Warn for when a key is too large for LMDB	2021-06-22 11:51:36 +02:00
Kerollmops	aecbd14761	Improve the error message for InvalidDocumentId	2021-06-22 11:31:58 +02:00
Kerollmops	0cca2ea24f	Return a MissingDocumentId when a document doesn't have one	2021-06-22 11:22:33 +02:00
Kerollmops	481b0bf277	Warn for when a facet key is too large for LMDB	2021-06-22 10:57:46 +02:00
bors[bot]	b073fd49ea	Merge #244 244: Update version for the next release (v0.5.0) r=Kerollmops a=curquiza Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2021-06-21 14:27:10 +00:00
bors[bot]	be2ebdd395	Merge #243 243: Rename FieldsDistribution into FieldDistribution r=Kerollmops a=curquiza Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2021-06-21 14:00:35 +00:00
Clémentine Urquizar	320670f8fe	Update version for the next release (v0.5.0)	2021-06-21 15:59:17 +02:00
Clémentine Urquizar	daef43f504	Rename FieldsDistribution into FieldDistribution	2021-06-21 15:57:41 +02:00
bors[bot]	b120c32cad	Merge #242 242: Update version for the next release (v0.4.2) r=Kerollmops a=curquiza Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2021-06-21 09:01:42 +00:00
Clémentine Urquizar	35fcc351a0	Update version for the next release (v0.4.2)	2021-06-20 17:37:24 +02:00
bors[bot]	5b19dd23d9	Merge #240 240: Field distribution r=Kerollmops a=irevoire closes #199 closes #198 Co-authored-by: Tamo <tamo@meilisearch.com>	2021-06-19 10:14:25 +00:00
Tamo	d08cfda796	convert the field_distribution to a BTreeMap and avoid counting twice the same documents	2021-06-17 18:31:54 +02:00
bors[bot]	a9e552ab18	Merge #238 238: Integration tests on filters and distinct r=Kerollmops a=ManyTheFish Fix #216 Fix #120 Co-authored-by: many <maxime@meilisearch.com>	2021-06-17 15:00:51 +00:00
many	6cb1102bdb	Fix PR comments	2021-06-17 15:19:03 +02:00
Tamo	969adaefdf	rename fields_distribution in field_distribution	2021-06-17 15:16:20 +02:00
bors[bot]	a67ccfdf3a	Merge #239 239: Update version to the next release (0.4.1) r=Kerollmops a=Kerollmops Co-authored-by: Kerollmops <clement@meilisearch.com>	2021-06-17 13:02:37 +00:00
Kerollmops	ccd6f13793	Update version to the next release (0.4.1)	2021-06-17 15:01:20 +02:00
many	f496cd320d	Add distinct integration tests	2021-06-17 14:33:18 +02:00
many	9f4184208e	Add test on filters	2021-06-17 13:56:09 +02:00
bors[bot]	bb89ef9fc0	Merge #237 237: change sub errors visibility r=Kerollmops a=MarinPostma re-export sub-error types so they can be matched upon outside of milli. Co-authored-by: marin postma <postma.marin@protonmail.com> Co-authored-by: marin <postma.marin@protonmail.com>	2021-06-17 09:51:18 +00:00
marin	70bee7d405	re-export remaining error types Co-authored-by: Clément Renault <clement@meilisearch.com>	2021-06-17 11:49:03 +02:00
marin postma	abbebad669	change sub errors visibility	2021-06-17 11:44:01 +02:00
bors[bot]	1bcf43baac	Merge #236 236: Format the whole project r=Kerollmops a=irevoire I need to add `cargo fmt` in the CI before closing #231 Co-authored-by: Tamo <tamo@meilisearch.com>	2021-06-16 18:05:40 +00:00
Tamo	9716fb3b36	format the whole project	2021-06-16 18:33:33 +02:00
bors[bot]	ba30cef987	Merge #234 234: Revert "Enable optimization in every profile" r=Kerollmops a=ManyTheFish compiling tests in release takes too much time. Reverts meilisearch/milli#224 Fix #233 Co-authored-by: Many <legendre.maxime.isn@gmail.com>	2021-06-16 13:38:58 +00:00
Many	41bdc90f46	Revert "Enable optimization in every profile"	2021-06-16 14:17:02 +02:00
bors[bot]	3bd4cf94cc	Merge #235 235: Update version for the next release (v0.4.0) r=Kerollmops a=curquiza Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2021-06-16 12:02:40 +00:00
Clémentine Urquizar	f5ff3e8e19	Update version for the next release (v0.4.0)	2021-06-16 14:01:05 +02:00
bors[bot]	02e0271e44	Merge #225 225: Introduce the error handler r=ManyTheFish a=Kerollmops Fixes #109. Co-authored-by: Kerollmops <clement@meilisearch.com> Co-authored-by: many <maxime@meilisearch.com>	2021-06-16 09:46:23 +00:00
many	ce0315a10f	Close write transaction in test	2021-06-16 11:03:37 +02:00
Kerollmops	7ac441e473	Fix small typos	2021-06-16 11:03:37 +02:00
Kerollmops	adf0c389c5	Rename FilterParsing into InvalidFilter	2021-06-16 11:03:36 +02:00
Kerollmops	8cfe3e1ec0	Rename DatabaseSizeReached into MaxDatabaseSizeReached	2021-06-16 11:03:36 +02:00
Kerollmops	4eda438f6f	Add a new Error for when a user use a non-filtered attribute in a filter	2021-06-16 11:03:36 +02:00
Kerollmops	713acc408b	Introduce the primary key to the Settings builder structure	2021-06-16 11:03:36 +02:00
Kerollmops	a7d6930905	Replace the panicking expect by tracked Errors	2021-06-15 11:51:32 +02:00
Kerollmops	f0e804afd5	Rename the FieldIdMapMissingEntry from_db_name field into process	2021-06-15 11:13:04 +02:00
Kerollmops	28c004aa2c	Prefer using constant for the database names	2021-06-15 11:13:04 +02:00
Kerollmops	78fe4259a9	Fix the http-ui crate	2021-06-14 18:06:23 +02:00
Kerollmops	312c2d1d8e	Use the Error enum everywhere in the project	2021-06-14 16:58:38 +02:00
Kerollmops	ca78cb5aca	Introduce more variants to the error module enums	2021-06-14 16:58:38 +02:00
Kerollmops	456541e921	Implement the Display trait on the Error type	2021-06-14 16:48:51 +02:00
Kerollmops	44c353fafd	Introduce some way to construct an Error	2021-06-14 16:48:51 +02:00
Kerollmops	23fcf7920e	Introduce a basic version of the InternalError struct	2021-06-14 16:48:51 +02:00
Kerollmops	d2b1ecc885	Remove a lot of serialization unreachable errors	2021-06-14 16:48:51 +02:00
Kerollmops	65b1d09d55	Move the obkv merging functions into the merge_function module	2021-06-14 16:48:51 +02:00
Kerollmops	ab727e428b	Remove the docid_word_positions_merge method that must never be called	2021-06-14 16:48:51 +02:00
Kerollmops	93a8633f18	Remove the documents_merge method that must never be called	2021-06-14 16:48:51 +02:00
Kerollmops	cfc7314bd1	Prefer using an explicit merge function name	2021-06-14 16:48:50 +02:00
Kerollmops	93978ec38a	Serializing a RoaringBitmap into a Vec cannot fail	2021-06-14 16:48:50 +02:00
Kerollmops	ff9414a6ba	Use the out of the compute_primary_key_pair function	2021-06-14 16:48:50 +02:00
bors[bot]	0542e2179f	Merge #230 230: Update Tokenizer version to v0.2.3 r=ManyTheFish a=curquiza Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2021-06-10 16:28:05 +00:00
Clémentine Urquizar	7d5395c12b	Update Tokenizer version to v0.2.3	2021-06-10 17:00:04 +02:00
bors[bot]	3e6c05fe13	Merge #227 227: Replace Consecutive by Phrase in query tree r=Kerollmops a=ManyTheFish Replace `Consecutive` by `Phrase` in the query tree in order to remove theoretical bugs, due to the `Consecutive` enum type. Co-authored-by: many <maxime@meilisearch.com> Co-authored-by: Many <legendre.maxime.isn@gmail.com>	2021-06-10 09:31:39 +00:00
Many	f4cab080a6	Update milli/src/search/query_tree.rs Co-authored-by: Clément Renault <clement@meilisearch.com>	2021-06-10 11:30:51 +02:00
Many	36715f571c	Update milli/src/search/criteria/proximity.rs Co-authored-by: Clément Renault <clement@meilisearch.com>	2021-06-10 11:30:33 +02:00
many	e923a3ed6a	Replace Consecutive by Phrase in query tree Replace Consecutive by Phrase in query tree in order to remove theorical bugs, due of the Consecutive enum type.	2021-06-10 11:16:16 +02:00
bors[bot]	bc02031793	Merge #226 226: Update version for the next release (v0.3.1) r=Kerollmops a=curquiza Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2021-06-09 14:13:42 +00:00
Clémentine Urquizar	dc64e139b9	Update version for the next release (v0.3.1)	2021-06-09 14:39:21 +02:00
bors[bot]	5cf1b0b138	Merge #224 224: Enable optimization in every profile r=Kerollmops a=irevoire Co-authored-by: Irevoire <tamo@meilisearch.com>	2021-06-09 09:13:21 +00:00
bors[bot]	afb4133bd2	Merge #212 #222 #223 212: Introduce integration test on criteria r=Kerollmops a=ManyTheFish - add pre-ranked dataset - test each criterion 1 by 1 - test all criteria in several order 222: Move the `UpdateStore` into the http-ui crate r=Kerollmops a=Kerollmops We no more need to have the `UpdateStore` inside of the mill crate as this is the job of the caller to stack the updates and sequentially give them to milli. 223: Update dataset links r=Kerollmops a=curquiza Co-authored-by: many <maxime@meilisearch.com> Co-authored-by: Many <legendre.maxime.isn@gmail.com> Co-authored-by: Kerollmops <clement@meilisearch.com> Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2021-06-09 08:47:19 +00:00
Irevoire	86b916b008	enable optimization in every profile	2021-06-09 10:26:57 +02:00
bors[bot]	6faa87302c	Merge #220 220: Make hard separators split phrase query r=Kerollmops a=ManyTheFish hard separators will now split a phrase query as two sequential phrases (double-quoted strings): the query `"Radioactive (Imagine Dragons)"` would be considered equivalent to `"Radioactive" "Imagine Dragons"` which as the little disadvantage of not keeping the order of the two (or more) separate phrases. Fix #208 Co-authored-by: many <maxime@meilisearch.com> Co-authored-by: Many <legendre.maxime.isn@gmail.com>	2021-06-09 08:22:58 +00:00
Many	f4ff30e99d	Update milli/tests/search/mod.rs Co-authored-by: Clément Renault <clement@meilisearch.com>	2021-06-09 10:12:24 +02:00
Many	ab696f6a23	Update milli/tests/search/query_criteria.rs Co-authored-by: Clément Renault <clement@meilisearch.com>	2021-06-09 10:12:17 +02:00
Clément Renault	d89f5ca48e	Merge pull request #219 from meilisearch/fix-criteria-fields-ids-map Save the criteria field name in the fields ids map	2021-06-08 18:46:57 +02:00
Clémentine Urquizar	7e93811fbc	Update dataset links	2021-06-08 18:18:54 +02:00
Kerollmops	0bf4f3f48a	Modify a test to check that criteria additions change the fields ids map	2021-06-08 18:14:34 +02:00
Kerollmops	82df524e09	Make sure that we register the field when setting criteria	2021-06-08 18:14:33 +02:00
Clément Renault	8e2c41e7f7	Merge pull request #221 from meilisearch/fix-primary-key-delete Use the index primary key when deleting documents	2021-06-08 18:13:42 +02:00
Kerollmops	103dddba2f	Move the UpdateStore into the http-ui crate	2021-06-08 17:59:51 +02:00
Many	faf148d297	Update milli/src/search/query_tree.rs Co-authored-by: Clément Renault <clement@meilisearch.com>	2021-06-08 17:52:37 +02:00
Kerollmops	133ab98260	Use the index primary key when deleting documents	2021-06-08 17:33:29 +02:00
many	b489d699ce	Make hard separators split phrase query hard separators will now split a phrase query as double double-quotes Fix #208	2021-06-08 17:29:38 +02:00
Many	afb09c914d	Update milli/tests/search/query_criteria.rs Co-authored-by: Clément Renault <clement@meilisearch.com>	2021-06-08 16:53:56 +02:00
many	b64cd2a3e3	Resolve PR comments	2021-06-08 14:14:34 +02:00
many	1fcc5f73ac	Factorize tests using macro_rules	2021-06-08 12:33:02 +02:00
bors[bot]	32cf5a29ce	Merge #218 218: Enable optimization for build.rs and macro r=Kerollmops a=irevoire It fasten the unzip of the benchmark’s dataset a lot Co-authored-by: Irevoire <tamo@meilisearch.com>	2021-06-08 09:56:23 +00:00
Irevoire	e0c327bae2	Update Cargo.toml Co-authored-by: Clément Renault <clement@meilisearch.com>	2021-06-08 11:39:10 +02:00
Irevoire	c82a382b0b	compile every build.rs with optimization	2021-06-08 11:19:22 +02:00
bors[bot]	eb149030eb	Merge #215 215: Make the benchmark command more convenient in CI r=Kerollmops a=curquiza Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2021-06-08 09:04:26 +00:00
bors[bot]	fd032165d7	Merge #217 217: Improve the benchmarks readme r=Kerollmops a=irevoire - Move the Dataset part to the end of the readme so when peoples just want to run the benchmarks they are not tempted to download the benchmarks by hand (which are going to be downloaded anyway by the `build.rs` scritp) - Fix the links in the dataset -- wiki part Co-authored-by: Irevoire <tamo@meilisearch.com>	2021-06-08 08:44:16 +00:00
Irevoire	d912c94034	improve the benchmark’s readme	2021-06-08 10:38:23 +02:00
Irevoire	563492f1e5	update the TOC order	2021-06-07 17:29:22 +02:00
Clémentine Urquizar	38ab541f4a	Make the benchmark command more convenient in CI	2021-06-04 00:21:39 +02:00
bors[bot]	af38196a6b	Merge #214 214: Add --locked in CI tests r=Kerollmops a=curquiza Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2021-06-03 14:39:36 +00:00
Clémentine Urquizar	e9104a0a32	Add --locked in CI tests	2021-06-03 16:23:59 +02:00
Clémentine Urquizar	70229f07c8	Update Cargo.lock	2021-06-03 16:22:43 +02:00
bors[bot]	ee7d291442	Merge #213 213: Fix the benchmarks script and names r=Kerollmops a=Kerollmops The benchmarks compare script was not using the `--output` flag and was therefore failing the download of the JSON reports. We also modified the criterion benchmarks to use shorter names, it helps in looking at the benchmarks in the terminal. Co-authored-by: Kerollmops <clement@meilisearch.com>	2021-06-03 14:18:45 +00:00
Kerollmops	29824d05ab	Reduce the length of the benchmarks names	2021-06-03 15:59:43 +02:00
Kerollmops	76a2343639	Fix the compare script of the benchmarks	2021-06-03 15:39:52 +02:00
many	10882bcbce	Introduce integration test on criteria	2021-06-03 14:44:53 +02:00
bors[bot]	a32236c80c	Merge #211 211: Update Cargo.toml for next release v0.3.0 r=Kerollmops a=curquiza Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2021-06-03 10:42:52 +00:00
Clémentine Urquizar	3b2b3aeea9	Update Cargo.toml for next release v0.3.0	2021-06-03 12:24:27 +02:00
bors[bot]	39ed133f9f	Merge #193 193: Fix primary key behavior r=Kerollmops a=MarinPostma this pr: - Adds early returns on empty document additions, avoiding error messages to be returned when adding no documents and no primary key was set. - Changes the primary key inference logic to match that of legacy meilisearch. close #194 Co-authored-by: Marin Postma <postma.marin@protonmail.com> Co-authored-by: marin postma <postma.marin@protonmail.com>	2021-06-03 10:24:21 +00:00
bors[bot]	fd598f060c	Merge #210 210: Check the benchmarks in the CI r=Kerollmops a=Kerollmops Fixes #209. Co-authored-by: Kerollmops <clement@meilisearch.com>	2021-06-03 09:16:06 +00:00
Kerollmops	99b45d2aa0	Make sure that all the workspaces crates compile	2021-06-03 10:56:01 +02:00
marin postma	57898d8a90	fix silent deserialize error	2021-06-03 10:42:55 +02:00
Kerollmops	82fb5f0bef	Fix the benchmarks compilation	2021-06-03 10:33:42 +02:00
Kerollmops	6b7841fefc	Make sure that the benchmarks always compile	2021-06-03 10:29:21 +02:00
bors[bot]	834504aec0	Merge #204 204: Decorrelate Distinct, Asc/Desc, Filterable fields from the faceted fields r=Kerollmops a=Kerollmops This PR decorrelates the fields that need to be stored in facet databases (big inverted indexes for fast access) from the filterable fields, the previously named faceted fields are now named filterable fields and are the union of the distinct attribute, all the Asc/Desc criteria and, the filterable fields. I added two tests to make sure that the engine was correctly generating the faceted databases when a distinct attribute or an Asc/Desc criteria were added, and one to make sure that it was impossible to filter on a non-filterable field even if it was a faceted one. Note that the `AttributesForFacetting` has also been renamed into `FilterableAttributes`. But it will be the Transplant's job to do that on the API, this change is only visible to the milli's library users. - Related to https://github.com/meilisearch/transplant/issues/187. - Fixes #161 by returning the documents that don't have the Asc/Desc field at the end of the bucket. - Fixes #168. - Fixes #152. Co-authored-by: Kerollmops <clement@meilisearch.com> Co-authored-by: Marin Postma <postma.marin@protonmail.com> Co-authored-by: many <maxime@meilisearch.com>	2021-06-02 15:43:39 +00:00
many	26a9974667	Make asc/desc criterion return resting documents Fix #161.2	2021-06-02 17:41:48 +02:00
bors[bot]	28962bce99	Merge #207 207: Benchmarks r=Kerollmops a=irevoire Co-authored-by: tamo <tamo@meilisearch.com> Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com> Co-authored-by: Tamo <irevoire@hotmail.fr> Co-authored-by: Irevoire <tamo@meilisearch.com>	2021-06-02 15:29:09 +00:00
Tamo	6dc08bf45e	remove the nop function	2021-06-02 17:09:21 +02:00
Tamo	087ae64899	add a gitignore to avoid pushing the autogenerated file	2021-06-02 17:03:30 +02:00
Tamo	3db25153e5	fix the faceted_fields one last time	2021-06-02 17:00:58 +02:00
Kerollmops	3c304c89d4	Make sure that we generate the faceted database when required	2021-06-02 16:24:58 +02:00
Kerollmops	b0c0490e85	Make sure that we can add a Asc/Desc field without it being filterable	2021-06-02 16:24:58 +02:00
Kerollmops	3b1cd4c4b4	Rename the FacetCondition into FilterCondition	2021-06-02 16:24:58 +02:00
Kerollmops	c2afdbb1fb	Move and comment some internal facet_condition helper functions	2021-06-02 16:24:58 +02:00
Kerollmops	6476827d3a	Fix the indexer to be sure that distinct and Asc/Desc are also faceted	2021-06-02 16:24:58 +02:00
Kerollmops	c10469ddb6	Patch the http-ui crate to support filterable fields	2021-06-02 16:24:58 +02:00
Marin Postma	1e366dae3e	remove useless lifetime on Distinct Trait	2021-06-02 16:24:58 +02:00
Kerollmops	187c713de5	Remove the MapDistinct struct as now distinct attributes are faceted	2021-06-02 16:24:57 +02:00
Kerollmops	ff440c1d9d	Introduce the faceted fields method to retrieve those that needs faceting	2021-06-02 16:24:57 +02:00
Kerollmops	2a3f9b32ff	Rename the faceted fields into filterable fields	2021-06-02 16:24:57 +02:00
Irevoire	f346805c0c	Update benchmarks/Cargo.toml Co-authored-by: Clément Renault <clement@meilisearch.com>	2021-06-02 15:47:03 +02:00
Clémentine Urquizar	ef1ac8a0cb	Update README	2021-06-02 11:13:22 +02:00
Clémentine Urquizar	edfcdb171c	Update benchmarks/scripts/list.sh Co-authored-by: Irevoire <tamo@meilisearch.com>	2021-06-02 11:13:22 +02:00
Clémentine Urquizar	3c91a9a551	Update following reviews	2021-06-02 11:13:22 +02:00
Tamo	bc4f4ee829	remove s3cmd as a dependency and provide a script to list all the available benchmarks	2021-06-02 11:13:22 +02:00
Clémentine Urquizar	61fe422a88	Update benchmarks/scripts/compare.sh Co-authored-by: Irevoire <tamo@meilisearch.com>	2021-06-02 11:13:22 +02:00
Clémentine Urquizar	57ed96622b	Update benchmarks/scripts/compare.sh Co-authored-by: Irevoire <tamo@meilisearch.com>	2021-06-02 11:13:22 +02:00
Clémentine Urquizar	b3c0d43890	Update benchmarks/scripts/compare.sh Co-authored-by: Irevoire <tamo@meilisearch.com>	2021-06-02 11:13:22 +02:00
Clémentine Urquizar	0d0e900158	Add CI for benchmarks	2021-06-02 11:13:22 +02:00
tamo	4536dfccd0	add a way to provide primary_key or autogenerate documents ids	2021-06-02 11:13:20 +02:00
tamo	06c414a753	move the benchmarks to another crate so we can download the datasets automatically without adding overhead to the build of milli	2021-06-02 11:11:50 +02:00
tamo	3c84075d2d	uses an env variable to find the datasets	2021-06-02 11:05:07 +02:00
tamo	4969abeaab	update the facets for the benchmarks	2021-06-02 11:05:07 +02:00
tamo	e5dfde88fd	fix the facets conditions	2021-06-02 11:05:07 +02:00
tamo	7c7fba4e57	remove the time limitation to let criterion do what it wants	2021-06-02 11:05:07 +02:00
tamo	5d5d115608	reformat all the files	2021-06-02 11:05:07 +02:00
tamo	7086009f93	improve the base search	2021-06-02 11:05:07 +02:00
tamo	d0b44c380f	add benchmarks on a wiki dataset	2021-06-02 11:05:07 +02:00
tamo	beae843766	add a missing space	2021-06-02 11:05:07 +02:00
tamo	5132a106a1	refactorize everything related to the songs dataset in a songs benchmark file	2021-06-02 11:05:07 +02:00
tamo	136efd6b53	fix the benches	2021-06-02 11:05:07 +02:00
tamo	4b78ef31b6	add the configuration of the searchable fields and displayed fields and a default configuration for the songs	2021-06-02 11:05:07 +02:00
tamo	ea0c6d8c40	add a bunch of queries and start the introduction of the filters and the new dataset	2021-06-02 11:05:07 +02:00
tamo	3def42abd8	merge all the criterion only benchmarks in one file	2021-06-02 11:05:07 +02:00
tamo	a2bff68c1a	remove the optional words for the typo criterion	2021-06-02 11:05:07 +02:00
tamo	aee49bb3cd	add the proximity criterion	2021-06-02 11:05:07 +02:00
tamo	49e4cc3daf	add the words criterion to the bench	2021-06-02 11:05:07 +02:00
tamo	15cce89a45	update the README with instructions to get the download the dataset	2021-06-02 11:05:07 +02:00
tamo	e425f70ef9	let criterion decide how much iteration it wants to do in 10s	2021-06-02 11:05:07 +02:00
tamo	4fdbfd6048	push a first version of the benchmark for the typo	2021-06-02 11:05:07 +02:00
bors[bot]	270da98c46	Merge #202 202: Add field id word count docids database r=Kerollmops a=LegendreM This PR introduces a new database, `field_id_word_count_docids`, that maps the number of words in an attribute with a list of document ids. This relation is limited to attributes that contain less than 11 words. This database is used by the exactness criterion to know if a document has an attribute that contains exactly the query without any additional word. Fix #165 Fix #196 Related to [specifications:#36](https://github.com/meilisearch/specifications/pull/36) Co-authored-by: many <maxime@meilisearch.com> Co-authored-by: Many <legendre.maxime.isn@gmail.com>	2021-06-01 16:09:48 +00:00
many	e857ca4d7d	Fix PR comments	2021-06-01 18:06:46 +02:00
Many	ab2cf69e8d	Update milli/src/update/delete_documents.rs Co-authored-by: Clément Renault <clement@meilisearch.com>	2021-06-01 17:04:10 +02:00
Many	8e6d1ff0dc	Update milli/src/update/index_documents/store.rs Co-authored-by: Clément Renault <clement@meilisearch.com>	2021-06-01 17:04:02 +02:00
bors[bot]	168fe0aa28	Merge #206 206: Fix http-ui r=Kerollmops a=irevoire I just noticed that `http-ui` was not compiling on `main`. I'm not sure this is the best fix, but it works 👀 Co-authored-by: Tamo <irevoire@hotmail.fr>	2021-06-01 14:31:32 +00:00
Tamo	608c5bad24	fix http-ui	2021-06-01 16:24:46 +02:00
bors[bot]	7d36d664a7	Merge #203 203: Make the MatchingWords return the number of matching bytes r=Kerollmops a=LegendreM Make the MatchingWords return the number of matching bytes using a custom Levenshtein algorithm. Fix #138 Co-authored-by: many <maxime@meilisearch.com>	2021-06-01 12:00:33 +00:00
many	225ae6fd25	Resolve PR comments	2021-06-01 11:53:09 +02:00
bors[bot]	2f9f6a1f21	Merge #169 169: Optimize roaring codec r=Kerollmops a=MarinPostma Optimize the `BoRoaringBitmapCodec` by preventing it from emiting useless error that caused allocation. On my flamegraph, the byte_decode function went from 4.13% to 1.70% (of transplant graph). This may not be the greatest optimization ever, but hey, this was a low hanging fruit. before: ![image](https://user-images.githubusercontent.com/28804882/116241125-17018880-a754-11eb-9f9d-a67418d100e1.png) after: ![image](https://user-images.githubusercontent.com/28804882/116241167-21bc1d80-a754-11eb-9afc-d9d72727477c.png) Co-authored-by: Marin Postma <postma.marin@protonmail.com>	2021-06-01 06:30:25 +00:00
Marin Postma	984dc7c1ed	rewrite roaring codec without byteorder.	2021-05-31 22:15:39 +02:00
Marin Postma	1373637da1	optimize roaring codec	2021-05-31 22:15:35 +02:00
many	1df68d342a	Make the MatchingWords return the number of matching bytes	2021-05-31 18:22:29 +02:00
many	b8e6db0feb	Add database in infos crate	2021-05-31 16:29:27 +02:00
many	c701f8bf36	Use field id word count database in exactness criterion	2021-05-31 16:27:28 +02:00
many	4ddf008be2	add field id word count database	2021-05-31 16:27:28 +02:00
bors[bot]	2f5e61bacb	Merge #184 184: Transfer numbers and strings facets into the appropriate facet databases r=Kerollmops a=Kerollmops This pull request is related to https://github.com/meilisearch/milli/issues/152 and changes the layout of the facets values, numbers and strings are now in dedicated databases and the user no more needs to define the type of the fields. No more conversion between the two types is done, numbers (floats and integers converted to f64) go to the facet float database and strings go to the strings facet database. There is one related issue that I found regarding CSVs, the values in a CSV are always considered to be strings, [meilisearch/specifications#28](`d916b57d74/text/0028-indexing-csv.md`) fixes this issue by allowing the user to define the fields types using `:` in the "CSV Formatting Rules" section. All previous tests on facets have been modified to pass again and I have also done hand-driven tests with the 115m songs dataset. Everything seems to be good! Fixes #192. Co-authored-by: Clément Renault <clement@meilisearch.com> Co-authored-by: Kerollmops <clement@meilisearch.com>	2021-05-31 13:32:58 +00:00
Kerollmops	1c0a5cd136	Resolve code modification suggestions	2021-05-31 15:22:50 +02:00
bors[bot]	76b9178b16	Merge #200 200: Fix plane sweep algorithm r=Kerollmops a=LegendreM Fix plain sweep algorithm after creating some tests on proximity. Co-authored-by: many <maxime@meilisearch.com>	2021-05-26 11:36:24 +00:00
many	a5e98cf46d	Fix plane sweep algorithm	2021-05-25 18:21:55 +02:00
Kerollmops	5012cc3a32	Fix the http-ui crate to support split facet databases	2021-05-25 11:31:06 +02:00
Kerollmops	28bd9e183e	Fix the infos crate to support split facet databases	2021-05-25 11:31:06 +02:00
Clément Renault	3a4a150ef0	Fix the tests and remaining warnings	2021-05-25 11:31:06 +02:00
Clément Renault	02c655ff1a	Refine the facet distribution to use both databases	2021-05-25 11:30:00 +02:00
Clément Renault	79efded841	Refine the FacetCondition from_array constructor	2021-05-25 11:30:00 +02:00
Clément Renault	f7efde11d9	Refine the facet condition to use both facet databases	2021-05-25 11:30:00 +02:00
Clément Renault	e62b89a2ed	Make the facet distinct work with the new split facets	2021-05-25 11:30:00 +02:00
Clément Renault	bd7b285bae	Split the update side to use the number and the strings facet databases	2021-05-25 11:30:00 +02:00
Clément Renault	038e03a4e4	Use both facet databases in the FacetIter type	2021-05-25 11:30:00 +02:00
Clément Renault	597144b0b9	Use both number and string facet databases in the distinct system	2021-05-25 11:29:59 +02:00
Clément Renault	837c1041c7	Clear and delete the documents from the facet database	2021-05-25 11:28:36 +02:00
Clément Renault	a56c46b6f1	Explode the string and f64 facet databases into two	2021-05-25 11:28:36 +02:00
Clément Renault	df7a32e3d0	Move the creation date initialization into a function	2021-05-25 11:28:35 +02:00
bors[bot]	49bee2ebc5	Merge #190 190: Make bucket candidates optionals r=Kerollmops a=LegendreM Before the bucket candidates were the result of the facet filters or result of the query tree. They will now be only the result of the query tree, making the number of candidates more consistent between the same request with or without facet filters. Fix some clippy warnings. Fix #186 Co-authored-by: many <maxime@meilisearch.com>	2021-05-24 11:19:32 +00:00
many	a3944a7083	Introduce a filtered_candidates field	2021-05-11 11:37:40 +02:00
many	efba662ca6	Fix clippy warnings in cirteria	2021-05-10 10:27:18 +02:00
many	e923d51b8f	Make bucket candidates optionals	2021-05-10 10:27:04 +02:00
Marin Postma	eeb0c70ea2	meilisearch compatible primary key inference	2021-05-06 22:42:32 +02:00
Marin Postma	313c362461	early return on empty document addition	2021-05-06 18:14:16 +02:00
Many	c620626515	Merge pull request #188 from meilisearch/exactness-criterion Exactness criterion	2021-05-06 17:56:21 +02:00
Many	44b6843de7	Fix pull request reviews Update milli/src/fields_ids_map.rs Update milli/src/search/criteria/exactness.rs Update milli/src/search/criteria/mod.rs	2021-05-06 14:31:03 +02:00
many	c1ce4e4ca9	Introduce mocked ExactAttribute step in exactness criterion	2021-05-06 14:28:31 +02:00
many	a3f8686fbf	Introduce exactness criterion	2021-05-06 14:28:30 +02:00
bors[bot]	25f75d4d03	Merge #189 189: Update version for the next release (v0.2.1) r=Kerollmops a=curquiza Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2021-05-05 15:28:56 +00:00
bors[bot]	7e63e32960	Merge #187 187: Fix fields distribution after documents merge r=Kerollmops a=shekhirin Resolves https://github.com/meilisearch/milli/issues/174 The problem was with calculation of fields distribution before the merge in `output_from_sorter()`. So if you'd import two documents with the same primary key value, fields distribution will count it as two documents, while `output_from_sorter()` will merge these documents into one. --- ```console ➜ Downloads cat short_movies.json [ {"id":"47474","title":"The Serpent's Egg","poster":"https://image.tmdb.org/t/p/w500/n7z0doFkXHcvo8QQWHLFnkEPXRU.jpg","overview":"The Serpent's Egg follows a week in the life of Abel Rosenberg, an out-of-work American circus acrobat living in poverty-stricken Berlin following Germany's defeat in World War I.","release_date":246844800,"genres":["Thriller","Drama","Mystery"]}, {"id":"47474","title":"The Serpent's Egg","poster":"https://image.tmdb.org/t/p/w500/n7z0doFkXHcvo8QQWHLFnkEPXRU.jpg","overview":"The Serpent's Egg follows a week in the life of Abel Rosenberg, an out-of-work American circus acrobat living in poverty-stricken Berlin following Germany's defeat in World War I.","release_date":246844800,"genres":["Thriller","Drama","Mystery"]} ] ➜ Downloads curl -X POST -H "Content-Type: text/json" --data-binary @short_movies.json 127.0.0.1:7700/indexes/movies/documents {"updateId":0} ``` ## Before ```console ➜ Downloads curl -s 127.0.0.1:7700/indexes/movies/stats \| jq { "numberOfDocuments": 1, "isIndexing": false, "fieldsDistribution": { "release_date": 2, "poster": 2, "title": 2, "overview": 2, "genres": 2, "id": 2 } } ``` ## After ```console ➜ Downloads curl -s 127.0.0.1:7700/indexes/movies/stats \| jq { "numberOfDocuments": 1, "isIndexing": false, "fieldsDistribution": { "poster": 1, "release_date": 1, "title": 1, "genres": 1, "id": 1, "overview": 1 } } ``` Co-authored-by: Alexey Shekhirin <a.shekhirin@gmail.com>	2021-05-05 14:45:08 +00:00
Clémentine Urquizar	1e11578ef0	Update version for the next release (v0.2.1)	2021-05-05 14:57:34 +02:00
Alexey Shekhirin	f8d0f5265f	fix(update): fields distribution after documents merge	2021-05-04 22:12:20 +03:00
bors[bot]	1207a058d0	Merge #185 185: Provide an iterator over all the documents in a milli index r=Kerollmops a=irevoire Co-authored-by: tamo <tamo@meilisearch.com>	2021-05-04 14:04:16 +00:00
tamo	d61566787e	provide an iterator over all the documents in a milli index	2021-05-04 11:23:51 +02:00
bors[bot]	c08f4599f2	Merge #183 183: remove tests on main r=Kerollmops a=MarinPostma remove testing on main since we now use bors for merging. Co-authored-by: Marin Postma <postma.marin@protonmail.com>	2021-05-03 15:06:28 +00:00
Marin Postma	bb5823c775	remove tests on main	2021-05-03 15:21:20 +02:00
bors[bot]	792225eaff	Merge #182 182: Upgrade Milli version (v0.2.0) r=Kerollmops a=curquiza Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2021-05-03 13:00:16 +00:00
Clémentine Urquizar	a8680887d8	Upgrade Milli version (v0.2.0)	2021-05-03 14:50:47 +02:00
bors[bot]	5b93d6ab91	Merge #181 181: Upgrade Tokenizer version (v0.2.2) r=Kerollmops a=curquiza Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2021-05-03 11:03:25 +00:00
bors[bot]	5c762b71dd	Merge #177 177: Add bors r=Kerollmops a=curquiza Co-authored-by: Clémentine Urquizar <clementine@meilisearch.com>	2021-05-03 10:57:09 +00:00
Clémentine Urquizar	c30f17fafb	Add bors	2021-05-03 12:29:30 +02:00
Clémentine Urquizar	34e02aba42	Upgrade Tokenizer version (v0.2.2)	2021-05-03 10:55:55 +02:00
Clément Renault	03bb95539b	Merge pull request #180 from shekhirin/disable-autogenerated-doc-ids Disable autogenerate_docids by default	2021-05-01 12:22:13 +02:00
Alexey Shekhirin	d81c0e8bba	feat(update): disable autogenerate_docids by default	2021-04-30 21:41:34 +03:00
Clément Renault	c112877a4a	Merge pull request #178 from meilisearch/visible-document-nb make document addition number visible	2021-04-29 21:54:51 +02:00
Marin Postma	e8e32e0ba1	make document addition number visible	2021-04-29 20:05:07 +02:00
Clément Renault	b31f36d68c	Merge pull request #173 from meilisearch/enhance-distinct-attributes Remove excluded document in criteria iterations	2021-04-29 12:14:44 +02:00
many	ee09e50e7f	Remove excluded document in criteria iterations - pass excluded document to criteria to remove them in higher levels of the bucket-sort - merge already returned document with excluded documents to avoid duplicas Related to #125 and #112 Fix #170	2021-04-29 12:09:38 +02:00
Clément Renault	374c2782ad	Merge pull request #176 from yanns/patch-1 do not use echo that espaces newline	2021-04-29 10:50:15 +02:00
Yann Simon	566c4a53c5	do not use echo that espaces newline Fix https://github.com/meilisearch/milli/issues/175	2021-04-29 09:25:35 +02:00
Many	5b9524e1ba	Merge pull request #172 from meilisearch/optimize-proximity-criterion Optimize proximity criterion	2021-04-28 15:41:57 +02:00
many	31607bf9cd	Add a threshold on proximity when choosing between linear/set algorithm	2021-04-28 14:57:22 +02:00
Clément Renault	5a10de1b9f	Merge pull request #122 from meilisearch/attribute-criterion Introduce the Attribute criterion	2021-04-28 14:34:50 +02:00
many	3b7e6afb55	Make some refacto and add documentation	2021-04-28 13:53:27 +02:00
Many	0add4d735c	Update milli/src/search/criteria/attribute.rs Co-authored-by: Clément Renault <clement@meilisearch.com>	2021-04-27 17:40:34 +02:00
Many	3794ffc952	Update milli/src/search/criteria/attribute.rs Co-authored-by: Clément Renault <clement@meilisearch.com>	2021-04-27 17:39:23 +02:00
Many	329bd4a1bb	Update milli/src/search/criteria/attribute.rs Co-authored-by: Clément Renault <clement@meilisearch.com>	2021-04-27 17:39:03 +02:00
Many	3b1358b62f	Update milli/src/search/criteria/attribute.rs Co-authored-by: Clément Renault <clement@meilisearch.com>	2021-04-27 17:32:19 +02:00
Many	c862b1bc6b	Update milli/src/search/criteria/attribute.rs Co-authored-by: Clément Renault <clement@meilisearch.com>	2021-04-27 17:32:10 +02:00
Many	e92d137676	Update milli/src/search/criteria/attribute.rs Co-authored-by: Clément Renault <clement@meilisearch.com>	2021-04-27 17:31:42 +02:00
Many	b3d6c6a9a0	Update milli/src/search/criteria/attribute.rs Co-authored-by: Clément Renault <clement@meilisearch.com>	2021-04-27 17:31:13 +02:00
Many	498c2b298c	Update milli/src/search/criteria/attribute.rs Co-authored-by: Clément Renault <clement@meilisearch.com>	2021-04-27 17:30:02 +02:00
Many	0e4e6dfada	Update milli/src/search/criteria/proximity.rs Co-authored-by: Clément Renault <clement@meilisearch.com>	2021-04-27 17:29:52 +02:00
Many	47d780b8ce	Update milli/src/search/criteria/mod.rs Co-authored-by: Irevoire <tamo@meilisearch.com>	2021-04-27 14:39:53 +02:00
Many	0daa0e170a	Fix PR comments Co-authored-by: Clément Renault <clement@meilisearch.com>	2021-04-27 14:39:53 +02:00
many	0d7d3ce802	Update roaring package	2021-04-27 14:39:53 +02:00
many	71740805a7	Fix forgotten typo tests	2021-04-27 14:39:53 +02:00
many	e77291a6f3	Optimize Atrribute criterion on big requests	2021-04-27 14:39:53 +02:00
many	716c8e22b0	Add style and comments	2021-04-27 14:39:52 +02:00
many	f853790016	Use the LCM of 10 first numbers to compute attribute rank	2021-04-27 14:39:52 +02:00
many	2b036449be	Fix the return of equal candidates in different pages	2021-04-27 14:39:52 +02:00
many	0efa011e09	Make a small code clean-up	2021-04-27 14:39:52 +02:00
many	17c8c6f945	Make set algorithm return None when nothing can be returned	2021-04-27 14:39:52 +02:00
many	b3e2280bb9	Debug attribute criterion * debug folding when initializing iterators	2021-04-27 14:39:52 +02:00
many	1eee0029a8	Make attribute criterion typo/prefix tolerant	2021-04-27 14:39:52 +02:00
many	59f58c15f7	Implement attribute criterion * Implement WordLevelIterator * Implement QueryLevelIterator * Implement set algorithm based on iterators Not tested + Some TODO to fix	2021-04-27 14:39:52 +02:00
Clément Renault	361193099f	Reduce the amount of branches when query tree flattened	2021-04-27 14:39:52 +02:00
Kerollmops	7ff4a2a708	Display the number of entries in the infos crate	2021-04-27 14:39:52 +02:00
Kerollmops	1aad66bdaa	Compute stats about the word prefix level positions database in the infos crate	2021-04-27 14:39:52 +02:00
Kerollmops	e65bad16cc	Compute the words prefixes at the end of an update	2021-04-27 14:39:52 +02:00
many	ab92c814c3	Fix attributes score	2021-04-27 14:35:43 +02:00
Clément Renault	0ad9499b93	Fix an indexing bug in the words level positions	2021-04-27 14:35:43 +02:00
Clément Renault	7aa5753ed2	Make the attribute positions range bounds to be fixed	2021-04-27 14:35:43 +02:00
Clément Renault	658f316511	Introduce the Initial Criterion	2021-04-27 14:35:43 +02:00
Kerollmops	89ee2cf576	Introduce the TreeLevel struct	2021-04-27 14:25:35 +02:00
Kerollmops	bd1a371c62	Compute the WordsLevelPositions only once	2021-04-27 14:25:34 +02:00
Kerollmops	8bd4f5d93e	Compute the biggest values of the words_level_positions_docids	2021-04-27 14:25:34 +02:00
Kerollmops	f713828406	Implement the clear and delete documents for the word-level-positions database	2021-04-27 14:25:34 +02:00
Kerollmops	3069bf4f4a	Fix and improve the words-level-positions computation	2021-04-27 14:25:34 +02:00
Kerollmops	6b1b42b928	Introduce an infos wordsLevelPositionsDocids subcommand	2021-04-27 14:25:34 +02:00
Kerollmops	e8cc7f9cee	Expose a route in the http-ui to update the WordsLevelPositions	2021-04-27 14:25:34 +02:00
Kerollmops	3a25137ee4	Expose and use the WordsLevelPositions update	2021-04-27 14:25:34 +02:00
Kerollmops	c765f277a3	Introduce the WordsLevelPositions update	2021-04-27 14:25:34 +02:00
Kerollmops	9242f2f1d4	Store the first word positions levels	2021-04-27 14:25:34 +02:00
Kerollmops	b0a417f342	Introduce the word_level_position_docids Index database	2021-04-27 14:25:34 +02:00
many	75e7b1e3da	Implement test Context methods	2021-04-27 14:25:34 +02:00
many	4ff67ec2ee	Implement attribute criterion for small amounts of candidates	2021-04-27 14:25:34 +02:00
Kerollmops	0f4c0beffd	Introduce the Attribute criterion	2021-04-27 14:25:34 +02:00
Clément Renault	3bcc1c0560	Merge pull request #164 from meilisearch/clippy-fixes Make clippy happy	2021-04-21 13:32:29 +02:00
tamo	f8dee1b402	[makes clippy happy] search/criteria/proximity.rs	2021-04-21 12:36:45 +02:00
tamo	7fa3a1d23e	makes clippy happy http-ui	2021-04-21 12:36:45 +02:00
Clément Renault	28a8df2f0a	Merge pull request #160 from shekhirin/query-words-limit Support query words limit	2021-04-21 11:14:09 +02:00
Alexey Shekhirin	6fa00c61d2	feat(search): support words_limit	2021-04-20 12:22:04 +03:00
Clément Renault	726fcf015a	Merge pull request #146 from meilisearch/facet-float-integer-becomes-number Facet float-integer becomes facet number	2021-04-20 10:31:47 +02:00
Kerollmops	c9b2d3ae1a	Warn instead of returning an error when a conversion fails	2021-04-20 10:23:31 +02:00
Kerollmops	2aeef09316	Remove debug logs while iterating through the facet levels	2021-04-20 10:23:31 +02:00
Kerollmops	51767725b2	Simplify integer and float functions trait bounds	2021-04-20 10:23:31 +02:00
Kerollmops	efbfa81fa7	Merge the Float and Integer enum variant into the Number one	2021-04-20 10:23:30 +02:00
Clément Renault	f5ec14c54c	Merge pull request #163 from meilisearch/next-release-v0.1.1 Update version for the next release (v0.1.1)	2021-04-19 15:52:13 +02:00
Clémentine Urquizar	127d3d028e	Update version for the next release (v0.1.1)	2021-04-19 14:48:13 +02:00
Clément Renault	1095874e7e	Merge pull request #158 from shekhirin/synonyms Support synonyms	2021-04-18 11:00:13 +02:00
Alexey Shekhirin	33860bc3b7	test(update, settings): set & reset synonyms fixes after review more fixes after review	2021-04-18 11:24:17 +03:00
Alexey Shekhirin	e39aabbfe6	feat(search, update): synonyms	2021-04-18 11:24:17 +03:00
Clément Renault	995d1a07d4	Merge pull request #162 from michaelchiche/patch-1	2021-04-17 09:47:08 +02:00
Michael Chiche	f6b06d6e5d	typo: wrong command in example	2021-04-16 20:08:43 +02:00
Clément Renault	19b6620a92	Merge pull request #125 from meilisearch/distinct Implement distinct attribute	2021-04-15 16:33:49 +02:00
Marin Postma	9c4660d3d6	add tests	2021-04-15 16:25:56 +02:00
Marin Postma	75464a1baa	review fixes	2021-04-15 16:25:56 +02:00
Marin Postma	2f73fa55ae	add documentation	2021-04-15 16:25:55 +02:00
Marin Postma	45c45e11dd	implement distinct attribute distinct can return error facet distinct on numbers return distinct error review fixes make get_facet_value more generic fixes	2021-04-15 16:25:55 +02:00
Clément Renault	6e126c96a9	Merge pull request #159 from meilisearch/upd-tokenizer-v0.2.1 Update Tokenizer version to v0.2.1	2021-04-14 19:02:36 +02:00
Clémentine Urquizar	2c5c79d68e	Update Tokenizer version to v0.2.1	2021-04-14 18:54:04 +02:00
Clément Renault	c2df51aa95	Merge pull request #156 from meilisearch/stop-words Stop words	2021-04-14 17:33:06 +02:00
tamo	dcb00b2e54	test a new implementation of the stop_words	2021-04-12 18:35:33 +02:00
tamo	da036dcc3e	Revert "Integrate the stop_words in the querytree" This reverts commit `12fb509d84`. We revert this commit because it's causing the bug #150. The initial algorithm we implemented for the stop_words was: 1. remove the stop_words from the dataset 2. keep the stop_words in the query to see if we can generate new words by integrating typos or if the word was a prefix => This was causing the bug since, in the case of “The hobbit”, we were always looking for something starting with “t he” or “th e” instead of ignoring the word completely. For now we are going to fix the bug by completely ignoring the stop_words in the query. This could cause another problem were someone mistyped a normal word and ended up typing a stop_word. For example imagine someone searching for the music “Won't he do it”. If that person misplace one space and write “Won' the do it” then we will loose a part of the request. One fix would be to update our query tree to something like that: --------------------- OR OR TOLERANT hobbit # the first option is to ignore the stop_word AND CONSECUTIVE # the second option is to do as we are doing EXACT t # currently EXACT he TOLERANT hobbit --------------------- This would increase drastically the size of our query tree on request with a lot of stop_words. For example think of “The Lord Of The Rings”. For now whatsoever we decided we were going to ignore this problem and consider that it doesn't reduce too much the relevancy of the search to do that while it improves the performances.	2021-04-12 18:35:33 +02:00
Clément Renault	f9eab6e0de	Merge pull request #151 from meilisearch/release-drafter Add release drafter files	2021-04-12 10:25:52 +02:00
Clémentine Urquizar	6a128d4ec7	Add release drafter files	2021-04-12 10:18:39 +02:00
Clément Renault	5efe67f375	Merge pull request #154 from shekhirin/shekhirin/fix-settings-serde-tests test(http): fix and refactor settings assert_(ser\|de)_tokens	2021-04-11 10:52:38 +02:00
Alexey Shekhirin	3af8fa194c	test(http): combine settings assert_(ser\|de)_tokens into 1 test	2021-04-10 12:13:59 +03:00
Clément Renault	0d09c64dde	Merge pull request #148 from shekhirin/shekhirin/setting-enum refactor(http, update): introduce setting enum	2021-04-09 22:48:58 +02:00
Alexey Shekhirin	84c1dda39d	test(http): setting enum serialize/deserialize	2021-04-08 17:03:40 +03:00
Alexey Shekhirin	dc636d190d	refactor(http, update): introduce setting enum	2021-04-08 17:03:40 +03:00
Clément Renault	2bcdd8844c	Merge pull request #141 from meilisearch/reorganize-criterion reorganize criterion	2021-04-01 19:50:16 +02:00
tamo	0a4bde1f2f	update the default ordering of the criterion	2021-04-01 19:45:31 +02:00
Clément Renault	ee3f93c029	Merge pull request #136 from shekhirin/index-fields-ids-distribution-cache feat(index): store fields distribution in index	2021-04-01 18:36:21 +02:00
Alexey Shekhirin	2658c5c545	feat(index): update fields distribution in clear & delete operations fixes after review bump the version of the tokenizer implement a first version of the stop_words The front must provide a BTreeSet containing the stop words The stop_words are set at None if an empty Set is provided add the stop-words in the http-ui interface Use maplit in the test and remove all the useless drop(rtxn) at the end of all tests Integrate the stop_words in the querytree remove the stop_words from the querytree except if it was a prefix or a typo more fixes after review	2021-04-01 19:12:35 +03:00
Alexey Shekhirin	27c7ab6e00	feat(index): store fields distribution in index	2021-04-01 18:35:19 +03:00
Clément Renault	67e25f8724	Merge pull request #128 from meilisearch/stop-words Stop words	2021-04-01 14:02:37 +02:00
tamo	12fb509d84	Integrate the stop_words in the querytree remove the stop_words from the querytree except if it was a prefix or a typo	2021-04-01 13:57:55 +02:00
tamo	a2f46029c7	implement a first version of the stop_words The front must provide a BTreeSet containing the stop words The stop_words are set at None if an empty Set is provided add the stop-words in the http-ui interface Use maplit in the test and remove all the useless drop(rtxn) at the end of all tests	2021-04-01 13:57:55 +02:00
tamo	62a8f1d707	bump the version of the tokenizer	2021-04-01 13:49:22 +02:00
Clément Renault	56777af8e4	Merge pull request #135 from shekhirin/index-fields-ids-distribution feat(index): introduce fields_ids_distribution	2021-03-31 17:53:45 +02:00
Alexey Shekhirin	9205b640a4	feat(index): introduce fields_ids_distribution	2021-03-31 18:44:47 +03:00
Clément Renault	f2a786ecbf	Merge pull request #134 from meilisearch/improve_httpui add a button to display or show the facets	2021-03-31 17:07:04 +02:00
tamo	13ce0ebb87	stop requestings the facets if the user has hidden them	2021-03-31 16:27:32 +02:00
tamo	bcc131e866	add a button to display or hide the facets	2021-03-31 16:18:53 +02:00
Clément Renault	529c8f0eb1	Merge pull request #131 from shekhirin/criterion-asc-desc-regex fix(criterion): compile asc/desc regex only once	2021-03-30 15:18:21 +02:00
Alexey Shekhirin	2cb32edaa9	fix(criterion): compile asc/desc regex only once use once_cell instead of lazy_static reorder imports	2021-03-30 16:07:14 +03:00
Clément Renault	5a1d3609a9	Merge pull request #127 from shekhirin/main feat(search, criteria): const candidates threshold	2021-03-30 14:07:19 +02:00
Alexey Shekhirin	1e3f05db8f	use fixed number of candidates as a threshold	2021-03-30 11:57:10 +03:00
Alexey Shekhirin	a776ec9718	fix division	2021-03-29 19:16:58 +03:00
Alexey Shekhirin	522e79f2e0	feat(search, criteria): introduce a percentage threshold to the asc/desc	2021-03-29 19:08:31 +03:00
Clément Renault	9ad8b74111	Merge pull request #123 from irevoire/pin_tokenizer select a specific release of the tokenizer instead of using the latests git commit	2021-03-25 22:58:11 +01:00
tamo	73dcdb27f6	select a specific release of the tokenizer instead of using the latests git commit	2021-03-25 15:00:18 +01:00
Clément Renault	6b7cc0022b	Merge pull request #118 from meilisearch/fix-offset fix broken offset	2021-03-15 22:15:18 +01:00
mpostma	9c27183876	fix broken offset	2021-03-15 20:23:50 +01:00
Clément Renault	25f8789aa5	Merge pull request #117 from meilisearch/update-license Update LICENSE	2021-03-15 16:26:22 +01:00
Clémentine Urquizar	3455082458	Update LICENSE	2021-03-15 16:15:14 +01:00
Clément Renault	b7b23cd4a8	Merge pull request #116 from meilisearch/index-metadata add index metadata	2021-03-15 14:20:50 +01:00
mpostma	f0210453a6	add updated at on put primary key	2021-03-15 14:05:48 +01:00
mpostma	615fe095e1	update index updated at on index writes	2021-03-15 14:05:47 +01:00
mpostma	80d0f9c49d	methods to update index time metadata	2021-03-15 14:05:47 +01:00
Clément Renault	c9f9d39b54	Merge pull request #114 from meilisearch/github-ci-use-main Rename master into main in the Github CI	2021-03-11 20:46:06 +01:00
Kerollmops	0cc3132f5a	Rename master into main in the Github CI	2021-03-11 14:44:47 +01:00
Clément Renault	38b6e8decd	Merge pull request #106 from meilisearch/optimize-words-typo-criteria Optimize the words criterion	2021-03-10 11:28:46 +01:00
Kerollmops	d48008339e	Introduce two new optional_words and authorize_typos Search options	2021-03-10 11:16:30 +01:00
Kerollmops	54b97ed8e1	Update the fetcher comments	2021-03-10 10:56:26 +01:00
Kerollmops	d301859bbd	Introduce a special word_derivations function for Proximity	2021-03-10 10:42:53 +01:00
Kerollmops	facfb4b615	Fix the bucket candidates	2021-03-10 10:42:53 +01:00
Kerollmops	42fd7dea78	Remove the useless typo cache	2021-03-10 10:42:53 +01:00
many	62a70c300d	Optimize words criterion	2021-03-10 10:42:53 +01:00
Clément Renault	c53be51460	Merge pull request #105 from meilisearch/optimize-number-of-documents Optimize the number_of_documents function	2021-03-10 10:39:12 +01:00
Kerollmops	f51eb46c69	Use the RoaringBitmapLenCodec to retrieve the count of documents	2021-03-09 10:25:39 +01:00
Clément Renault	7a3ce9bb1d	Merge pull request #104 from meilisearch/update-license Update the LICENSE file to match the year 2021	2021-03-08 19:11:05 +01:00
Clément Renault	2f9af6a707	Fix the REAMD.md bash example	2021-03-08 18:56:22 +01:00
Clément Renault	f204344102	Update the LICENSE file to match the year 2021	2021-03-08 18:54:06 +01:00
Clément Renault	22f20f0c29	Merge pull request #99 from meilisearch/infos-missing-db-names Add missing databases to the infos subcommand	2021-03-08 18:52:08 +01:00
Clément Renault	18844d60b5	Simplify the output of database sizes in the infos crate	2021-03-08 18:47:33 +01:00
Clément Renault	3d02b19fbd	Introduce the docids-words-positions subcommand to the infos crate	2021-03-08 18:47:33 +01:00
Kerollmops	bd63da0a0e	Add missing databases to the infos subcommand	2021-03-08 18:47:33 +01:00
Clément Renault	f9be3ad3fd	Merge pull request #103 from meilisearch/plane-sweep-proximity Plane-Sweep proximity	2021-03-08 16:58:34 +01:00
Kerollmops	d781a6164a	Rewrite some code with idiomatic Rust	2021-03-08 16:27:52 +01:00
Clément Renault	b18ec00a7a	Add a logging_timer macro to te criterion next methods	2021-03-08 16:12:06 +01:00
Kerollmops	82a0f678fb	Introduce a cache on the docid_word_positions database method	2021-03-08 16:12:03 +01:00
Clément Renault	5fcaedb880	Introduce a WordDerivationsCache struct	2021-03-08 16:00:53 +01:00
many	2606c92ef9	use plain sweep in proximity criterion	2021-03-08 15:58:39 +01:00
many	ae47bb3594	Introduce plane_sweep function in proximity criterion	2021-03-08 15:58:38 +01:00
Kerollmops	636a9df177	Temporarily fix the tinytemplate doc hidden issue	2021-03-08 15:57:45 +01:00
Clément Renault	f190d5f496	Merge pull request #100 from meilisearch/improve-asc-desc-criterion Improve the Asc/Desc criteria	2021-03-08 13:37:00 +01:00
Clément Renault	3c76b3548d	Rework the Asc/Desc criteria to be facet iterator based	2021-03-08 13:32:25 +01:00
Clément Renault	a58d2b6137	Print the Asc/Desc criterion field name in the debug prints	2021-03-08 13:32:25 +01:00
Clément Renault	08a0ff7091	Merge pull request #101 from meilisearch/criterion-display implement display for criterion	2021-03-08 13:29:05 +01:00
mpostma	e3095be85c	Remove Debug use in Display impl	2021-03-08 12:09:09 +01:00
mpostma	9e1eb25232	implement display for criterion Update milli/src/criterion.rs Co-authored-by: Clément Renault <clement@meilisearch.com>	2021-03-08 11:00:30 +01:00
Clément Renault	71b069d3e1	Merge pull request #102 from meilisearch/fix-searchable-settings-test Fix the searchable settings test	2021-03-08 10:55:30 +01:00
Clément Renault	e5bb96bc3b	Fix the searchable settings test	2021-03-06 12:48:41 +01:00
Clément Renault	2924ed31f3	Merge pull request #97 from meilisearch/criteria Introduce all the criteria	2021-03-03 18:24:22 +01:00
Kerollmops	9b6b35d9b7	Clean up some comments	2021-03-03 18:19:10 +01:00
Kerollmops	2cc4a467a6	Change the criterion output that cannot fail	2021-03-03 18:18:33 +01:00
Kerollmops	1fc25148da	Remove useless where clauses for the criteria	2021-03-03 18:09:19 +01:00
Kerollmops	07784c8990	Tune the words prefixes threshold to compute for 1/1000 instead	2021-03-03 15:51:28 +01:00
Kerollmops	f376c6a728	Make sure we retrieve the docid word positions	2021-03-03 15:45:03 +01:00
Kerollmops	5c5e51095c	Fix the Asc/Desc criteria to alsways return the QueryTree when available	2021-03-03 15:45:03 +01:00
many	cdaa96df63	optimize proximity criterion	2021-03-03 15:45:03 +01:00
many	246286f0eb	take hard separator into account	2021-03-03 15:45:03 +01:00
Kerollmops	6bf6b40495	Remove unused files	2021-03-03 15:45:03 +01:00
Kerollmops	f118d7e067	build criteria from settings	2021-03-03 15:45:03 +01:00
Kerollmops	025835c5b2	Fix the criteria to avoid always returning a placeholder	2021-03-03 15:45:03 +01:00
Kerollmops	36c1f93ceb	Do an union of the bucket candidates	2021-03-03 15:45:03 +01:00
many	b0e0c5eba0	remove option of bucket_candidates	2021-03-03 15:45:03 +01:00
Kerollmops	daf126a638	Introduce the final Fetcher criterion	2021-03-03 15:45:03 +01:00
many	7ac09d7b7c	remove option of bucket_candidates	2021-03-03 15:45:03 +01:00
Kerollmops	5af63c74e0	Speed-up the MatchingWords highlighting struct	2021-03-03 15:45:03 +01:00
Kerollmops	4510bbccca	Add a lot of debug	2021-03-03 15:43:44 +01:00
Kerollmops	ae4a237e58	Fix the maximum_proximity function	2021-03-03 15:43:44 +01:00
Kerollmops	9bc9b36645	Introduce the Proximity criterion	2021-03-03 15:43:44 +01:00
Kerollmops	22b84fe543	Use the words criterion in the search module	2021-03-03 15:43:44 +01:00
many	3d731cc861	remove option on bucket_candidates	2021-03-03 15:43:44 +01:00
Clément Renault	14f9f85c4b	Introduce the AscDesc criterion	2021-03-03 15:43:44 +01:00
many	b5b7ec0162	implement initial state for words criterion	2021-03-03 15:43:44 +01:00
Kerollmops	3415812b06	Imrpove the intersection speed in the words criterion	2021-03-03 15:43:43 +01:00
Clément Renault	ef381e17bb	Compute the candidates for each sub query tree	2021-03-03 15:43:43 +01:00
Kerollmops	e174ccbd8e	Use the words criterion in the search module	2021-03-03 15:43:43 +01:00
Clément Renault	1e47f9b3ff	Introduce the Words criterion	2021-03-03 15:43:43 +01:00
many	2d068bd45b	implement Context trait for criteria	2021-03-03 15:43:43 +01:00
many	d92ad5640a	remove option on bucket_candidates	2021-03-03 15:43:43 +01:00
many	64688b3786	fix query tree builder	2021-03-03 15:43:43 +01:00
many	fb7e6df790	add tests on typo criterion	2021-03-03 15:43:43 +01:00
Kerollmops	c5a32fd4fa	Fix the typo criterion	2021-03-03 15:43:42 +01:00
many	a273c46559	clean warnings	2021-03-03 15:43:42 +01:00
many	9e093d5ff3	add cache on alterate_query_tree function	2021-03-03 15:43:42 +01:00
many	41fc51ebcf	optimize alterate_query_tree when number_typos is zero	2021-03-03 15:43:42 +01:00
many	4da6e1ea9c	add cache in typo criterion	2021-03-03 15:43:42 +01:00
Kerollmops	67c71130df	Reduce the number of calls to alterate_query_tree	2021-03-03 15:43:42 +01:00
many	9ccaea2afc	simplify criterion context	2021-03-03 15:43:42 +01:00
Clément Renault	fea9ffc46a	Use the bucket candidates in the search module	2021-03-03 15:43:42 +01:00
Clément Renault	229130ed25	Correctly compute the bucket candidates for the Typo criterion	2021-03-03 15:43:42 +01:00
Clément Renault	5344abc008	Introduce the CriterionResult return type	2021-03-03 15:43:41 +01:00
many	86bcecf840	change variable's name from distance to proximity	2021-03-03 15:43:41 +01:00
many	4128bdc859	reduce match possibilities in docids fetchers	2021-03-03 15:43:41 +01:00
many	907482c8ac	clean docids fetchers	2021-03-03 15:43:41 +01:00
many	774a255f2e	use prefix cache in criteria	2021-03-03 15:43:41 +01:00
many	98e69e63d2	implement Context trait for criteria	2021-03-03 15:43:41 +01:00
Clément Renault	f091f370d0	Use the Typo criteria in the search module	2021-03-03 15:43:41 +01:00
Clément Renault	ad20d72a39	Introduce the Typo criterion	2021-03-03 15:43:41 +01:00
Clément Renault	f0ddea821c	Introduce the Typo criterion	2021-03-03 15:43:41 +01:00
many	73286dc8bf	Introduce the query tree data structure	2021-03-03 15:43:40 +01:00
Clément Renault	4e84999f20	Merge pull request #80 from meilisearch/query_tree Introduce the query tree data structure	2021-03-03 14:25:29 +01:00
Kerollmops	411a118148	Avoid testing on nightly to fix a crate issue	2021-03-03 13:57:36 +01:00
Kerollmops	240b02e175	Remove unused Operation constructors	2021-03-03 13:40:19 +01:00
many	a463ae821e	Add methods optional_words and authorize_typos on the query tree	2021-03-03 13:40:19 +01:00
Kerollmops	6d135beb21	Introduce the maximum_proximity helper function	2021-03-03 13:40:18 +01:00
Kerollmops	6008f528d0	Introduce the maximum_typo helper function	2021-03-03 13:40:18 +01:00
Kerollmops	1dc857a4b2	Fix the query tree optional word generation with phrases	2021-03-03 13:40:18 +01:00
Kerollmops	4f19749252	Introduce the word_documents_count method on the Context trait	2021-03-03 13:40:18 +01:00
Kerollmops	79a143b32f	Introduce the query tree data structure	2021-03-03 13:40:18 +01:00
Clément Renault	5f109e8589	Merge pull request #95 from meilisearch/helpers-crate Introduce an helpers crate that export the database to stdout	2021-03-01 19:59:18 +01:00
Clément Renault	9423310816	Introduce an helpers crate that export the database to stdout	2021-03-01 19:55:04 +01:00
Clément Renault	68102fced8	Merge pull request #86 from meilisearch/clean-up-infos-crate Clean up the infos crate	2021-03-01 19:54:21 +01:00
Clément Renault	1eb7ce5cdb	Improve the export-documents infos command by accepting internal ids	2021-03-01 19:48:01 +01:00
Clément Renault	4884b324e6	Remove the useless external ids patch method in the infos crate	2021-03-01 19:48:01 +01:00
Clément Renault	78bede1ffb	Fix error displaying of the workspace members	2021-03-01 19:48:01 +01:00
Clément Renault	b59fe77ec7	Avoid creating a default empty database in the search crate	2021-03-01 19:48:01 +01:00
Clément Renault	45330a5e47	Avoid creating a default empty database in the infos crate	2021-03-01 19:48:00 +01:00
Clément Renault	794fce7bff	Merge pull request #91 from meilisearch/add-primary-key-to-fields-map add primary key to fields_id_map when not present	2021-03-01 16:20:41 +01:00
mpostma	e08b6b3ec7	add primary key to fields_id_map when not present	2021-03-01 16:10:16 +01:00
Clément Renault	8dcb3e0c41	Merge pull request #90 from meilisearch/words-prefixes-update Expose the WordsPrefixes update from the UpdateBuilder	2021-02-21 12:27:48 +01:00
Clément Renault	c62d2f56d8	Expose an http route for the WordsPrefixes update	2021-02-21 12:16:53 +01:00
Clément Renault	c318373b88	Expose the WordsPrefixes update on the UpdateBuilder	2021-02-21 12:15:35 +01:00
Clément Renault	3090751dfc	Merge pull request #94 from meilisearch/update-dependencies Update dependencies	2021-02-21 12:08:18 +01:00
Kerollmops	519b1cb5c9	Update dependencies	2021-02-21 10:26:04 +01:00
Clément Renault	e62157e896	Merge pull request #88 from meilisearch/heed-error-word-documents-count Return an heed error from the word_documents_count method	2021-02-18 15:05:00 +01:00
Kerollmops	c2ffcc4bd1	Return an heed error from the word_documents_count method	2021-02-18 14:59:37 +01:00
Clément Renault	09ca5d14c9	Merge pull request #87 from meilisearch/roaring-bitmap-length Introduce fast methods to get roaring bitmap lengths	2021-02-18 14:52:40 +01:00
Kerollmops	2f561c77f5	Introduce the word documents count method on the index	2021-02-18 14:35:14 +01:00
Kerollmops	8d710c5130	Introduce heed codecs to retrieve the length of roaring bitmaps	2021-02-18 14:30:47 +01:00
Kerollmops	fcfb39c5de	Move the RoaringBitmap related codecs into a module	2021-02-18 13:56:28 +01:00
Clément Renault	85c3d8aa52	Merge pull request #79 from meilisearch/prefix-caches Introduce prefix databases	2021-02-17 11:27:15 +01:00
Kerollmops	aa4d9882d2	Introduce the new words-prefixes-docids infos subcomand	2021-02-17 11:22:27 +01:00
Kerollmops	49aee6d02c	Fix the database-stats infos subcommand	2021-02-17 11:22:27 +01:00
Kerollmops	7a0f86a04f	Introduce an infos command to extract the words prefixes fst	2021-02-17 11:22:27 +01:00
Kerollmops	a4a48be923	Run the words prefixes update inside of the indexing documents update	2021-02-17 11:22:26 +01:00
Kerollmops	8788485924	Take the prefix databases into account in the infos subcommand	2021-02-17 11:22:26 +01:00
Kerollmops	616ed8f73c	Clean up the word prefix pair proximities when deleting documents	2021-02-17 11:22:26 +01:00
Clément Renault	ea37fd821d	Clean up the words prefixes when deleting documents and words	2021-02-17 11:22:25 +01:00
Clément Renault	62eee9c69e	Introduce the sorter_into_lmdb_database helper function	2021-02-17 11:12:39 +01:00
Clément Renault	b5b89990eb	Compute and write the word prefix pair proximities database	2021-02-17 11:12:38 +01:00
Kerollmops	9b03b0a1b2	Introduce the word prefix pair proximity docids database	2021-02-17 11:12:38 +01:00
Clément Renault	f365de636f	Compute and write the word-prefix-docids database	2021-02-17 11:12:38 +01:00
Clément Renault	ee5a60e1c5	Clear the words prefixes when clearing an index	2021-02-17 10:45:17 +01:00
Clément Renault	5e7b26791b	Take the words-prefixes into account while computing the biggest values	2021-02-17 10:45:17 +01:00
Clément Renault	b3a21d5a50	Introduce the getters and setters for the words prefixes FST	2021-02-17 10:45:17 +01:00
Clément Renault	48b470140b	Merge pull request #84 from meilisearch/stringify-documents-ids Stringify documents ids even when deleting documents	2021-02-15 21:30:51 +01:00
Clément Renault	89ce4e74fe	Do not change the primary key type when we serialize documents	2021-02-15 21:24:36 +01:00
Clément Renault	69acdd437e	Deserialize documents ids into JSON Values on deletion	2021-02-15 21:24:36 +01:00
Clément Renault	b3776598d8	Add a test to check deletion of documents with number as primary key	2021-02-15 21:24:35 +01:00
Clément Renault	5d0ac3e3e6	Merge pull request #81 from meilisearch/smart-workspace Change the project to become a workspace	2021-02-14 19:02:00 +01:00
Clément Renault	fecf3d6fc1	Move the command lines helpers into different crates	2021-02-14 18:55:15 +01:00
Clément Renault	d8f3421608	Update the dependencies and remove the unused ones	2021-02-14 18:32:46 +01:00
Clément Renault	e8639517da	Change the project to become a workspace with milli as a default-member	2021-02-12 16:15:09 +01:00
Clément Renault	d450b971f9	Merge pull request #78 from meilisearch/required-changes-for-transplant Changes for transplant	2021-02-02 16:22:09 +01:00
mpostma	8f43698a60	fix httpui	2021-02-01 19:49:51 +01:00
mpostma	3b60432687	Use update_id in UpdateBuilder Add `the update_id` to the to the updates. The rationale is the following: - It allows for better tracability of the update events, thus improved debugging and logging. - The enigne is now aware of what he's already processed, and can return it if asked. It may not make sense now, but in the future, the update store may not work the same way, and this information about the state of the engine will be desirable (distributed environement).	2021-02-01 19:46:34 +01:00
mpostma	d487791b03	derive serde for method and format This is nicer when working with UpdateMeta struct	2021-02-01 19:46:34 +01:00
mpostma	91d8198d17	return documents number on addition	2021-02-01 19:42:10 +01:00
Clément Renault	fa0cc2dc13	Merge pull request #66 from meilisearch/show-available-facets Expose an API to compute facets distribution	2021-02-01 18:39:45 +01:00
Clément Renault	14ae01a6c9	Fix some typos in error messages	2021-02-01 18:10:57 +01:00
Clément Renault	f5f4438b43	Remove the duplicated code inside the facet_values_from_documents method	2021-01-28 11:22:18 +01:00
Clément Renault	b6e91291fb	Add a comment to explain Serialize on FacetValue is implemented by hand	2021-01-27 18:29:56 +01:00
Clément Renault	b41bf58658	Split the FacetDistribution facet_values method into three	2021-01-27 18:29:56 +01:00
Clément Renault	a3e3bebed7	Rework the FacetDistribution execute method to use the faceted_fields struct	2021-01-27 18:29:54 +01:00
Clément Renault	11309ee99c	Rework the FacetDistribution execute method to use the faceted_fields struct	2021-01-27 14:53:50 +01:00
Clément Renault	9c8a654079	Add comments to help read the facet_values branchings	2021-01-27 14:49:08 +01:00
Clément Renault	2e00740515	Make sure that we don't iterate throught all string facet values	2021-01-27 14:41:36 +01:00
Clément Renault	b52d500fbc	Reorder the FacetType enum branching in the facet_value method	2021-01-27 14:36:49 +01:00
Clément Renault	d91d321129	Introduce some constants to the FacetDistribution struct and settings	2021-01-27 14:32:30 +01:00
Clément Renault	60480a1e2f	Rework the FacetCondition from_array constructor	2021-01-27 14:25:53 +01:00
Clément Renault	65b821b192	Rename the Index facets method into facets_distribution	2021-01-27 14:15:33 +01:00
Clément Renault	433ac8c38a	Remove the ordered-float serde feature	2021-01-27 14:11:10 +01:00
Clément Renault	70e9b1e936	Introduce a flag to the search subcommand to display the facet distribution	2021-01-26 14:58:18 +01:00
Kerollmops	61dbcfa44a	Bump the roaring to 0.6.4	2021-01-26 14:38:43 +01:00
Kerollmops	916dd3b7c5	Use the faceted_fields_ids method to fetch the ids	2021-01-26 14:14:38 +01:00
Clément Renault	b0c31500fc	Simplify the front page	2021-01-26 14:14:38 +01:00
Kerollmops	7be275b692	Add the count to the facet distribution	2021-01-26 14:14:37 +01:00
Clément Renault	4b9e81fc89	Order the facet values lexicographically	2021-01-26 14:09:09 +01:00
Clément Renault	51a37de885	Introduce the FacetValue enum type	2021-01-26 14:09:09 +01:00
Kerollmops	d893e83622	Speed-up facet aggregation by using a FacetIter	2021-01-26 14:09:08 +01:00
Kerollmops	33945a3115	Introduce a new facet filters query field	2021-01-26 14:09:08 +01:00
Kerollmops	afa86d8a45	Add a simple test to the FacetCondition from_array method	2021-01-26 14:06:29 +01:00
Kerollmops	cb5e57e2dd	FacetCondition can be created from array of facets	2021-01-26 14:06:28 +01:00
Clément Renault	a8e3269ad6	Introduce a basic front to display facets	2021-01-26 14:06:28 +01:00
Clément Renault	2cd8675734	Show facet values even for empty queries	2021-01-26 14:06:28 +01:00
Clément Renault	3916c54501	Speed-up facet aggregation on low number of candidates	2021-01-26 14:06:28 +01:00
Clément Renault	a17bb54d8f	Limit the number of values by facets to a maximum of 1000	2021-01-26 14:06:28 +01:00
Kerollmops	aa129dd7e8	Display the number of candidates instead of the returned document count	2021-01-26 14:06:28 +01:00
Kerollmops	510df4729c	Append the facet value to the facet query on click	2021-01-26 14:06:28 +01:00
Kerollmops	d25a859985	Display the facet values on the HTML debug page	2021-01-26 14:06:28 +01:00
Kerollmops	3b64735058	Introduce a struct to compute facets values	2021-01-26 14:06:27 +01:00
Clément Renault	30dae0205e	Merge pull request #67 from meilisearch/fix-settings Fix displayed and searchable attributes	2021-01-26 14:03:43 +01:00
mpostma	87a56d2bc9	Fix settings bug replace ids with str in settings This allows for better maintainability of the settings code, since updating the searchable attributes is now straightforward. criterion use string fix reindexing fieldid remaping add tests for primary_key compute fix tests fix http-ui fixup! add tests for primary_key compute code improvements settings update deps fixup! code improvements settings fixup! refactor settings updates and fix bug fixup! Fix settings bug fixup! Fix settings bug fixup! Fix settings bug Update src/update/index_documents/transform.rs Co-authored-by: Clément Renault <clement@meilisearch.com> fixup! Fix settings bug	2021-01-26 13:53:08 +01:00
Clément Renault	26f060f66b	Merge pull request #75 from meilisearch/fix-search-subcommand Fix the search subcommand document display loop	2021-01-20 10:07:16 +01:00
Clément Renault	c35befbf38	Fix the search subcommand document display loop	2021-01-18 19:06:36 +01:00
Clément Renault	2fa5808e3f	Merge pull request #71 from meilisearch/cleanup-useless-build-rs Cleanup useless custom build file	2021-01-15 15:45:47 +01:00
Clément Renault	44c0dd0762	Fix an fst Set related warning	2021-01-13 11:03:03 +01:00
Clément Renault	1bb9348a90	Remove the chinese-words.txt previous tokenizer related file	2021-01-13 11:01:57 +01:00
Clément Renault	9141f5ef94	Remove the custom build.rs file	2021-01-13 11:01:38 +01:00
Clément Renault	51d1785576	Merge pull request #63 from meilisearch/meilisearch-tokenizer Meilisearch tokenizer	2021-01-12 13:26:24 +01:00
mpostma	4f7f7538f7	highlight with new tokenizer	2021-01-11 21:59:37 +01:00
mpostma	1ae761311e	integrate with meilisearch tokenizer	2021-01-07 16:14:27 +01:00
Clément Renault	7e1c94ab9c	Merge pull request #65 from meilisearch/improve-facet-value-display Improve the facet value displaying	2021-01-07 16:12:32 +01:00
Clément Renault	0a1beb688c	Improve the facet value displaying, extracting the facet level	2021-01-07 16:05:09 +01:00
Clément Renault	5dd4dc2862	Merge pull request #60 from meilisearch/accept-compressed-documents-updates Accept and mirror compression of documents additions	2020-12-23 10:59:26 +01:00
Kerollmops	a576c7ae4b	Display the update meta result content on the update page	2020-12-22 13:42:43 +01:00
Kerollmops	6c7db3d956	Display the time it took to process an update	2020-12-22 13:42:43 +01:00
Kerollmops	9fcbc83ebc	Accept and mirror compression of documents additions	2020-12-22 13:42:42 +01:00
Clément Renault	cd158d4cde	Merge pull request #61 from meilisearch/update-handler create update handler trait	2020-12-22 13:42:00 +01:00
mpostma	49a016b53d	create update handler trait fix type inference error	2020-12-22 12:59:15 +01:00
Clément Renault	5039528b56	Merge pull request #59 from meilisearch/improve-bytes-structopt Use the byte-unit crate to ease library usage	2020-12-20 14:52:39 +01:00
Kerollmops	77e951e933	Use the byte-unit crate to ease library usage	2020-12-20 12:00:37 +01:00
Clément Renault	b032ceb5d4	Merge pull request #56 from meilisearch/asc-desc-criteria-non-faceted Return non-faceted documents to complete the requested limit	2020-12-17 14:34:36 +01:00
Clément Renault	914eab12f7	Return non-faceted documents as remaining results	2020-12-17 13:57:07 +01:00
Clément Renault	0dec761e21	Merge pull request #54 from meilisearch/compress-updates Compress updates content using gzip	2020-12-17 11:06:31 +01:00
Clément Renault	5a23417499	Compress updates content using gzip	2020-12-17 10:59:58 +01:00
Clément Renault	cd5605bb86	Merge pull request #50 from meilisearch/fix-asc-desc-criterion Fix the Asc/Desc criteria	2020-12-13 11:59:11 +01:00
Clément Renault	0e5609d40e	Limit the number of elements after reversing it	2020-12-12 14:21:27 +01:00
Clément Renault	9d966a28d3	Merge pull request #47 from meilisearch/fix-grenad-write-bug Bump grenad to fix an indexing bug	2020-12-05 17:18:49 +01:00
Clément Renault	e7f2ab9138	Bump grenad to fix an indexing bug	2020-12-05 16:39:15 +01:00
Clément Renault	9628da2d17	Merge pull request #40 from meilisearch/asc-desc-faceted-fields Ascending and descending custom ranking	2020-12-04 12:08:22 +01:00
Clément Renault	026f54dcf7	Use the field id docid facet value database when sorting documents	2020-12-04 12:03:20 +01:00
Clément Renault	3cdf14d4c5	Introduce the field-id-docid-facet-values database	2020-12-04 12:03:20 +01:00
Clément Renault	4ffbddf21f	Introduce debug info for the time it takes to fetch candidates	2020-12-04 12:03:20 +01:00
Clément Renault	13217f072b	Use the FacetRange iterator in the facet exploring function	2020-12-04 12:03:20 +01:00
Clément Renault	0959e1501f	Introduce the FacetRevRange Iterator struct	2020-12-04 12:02:23 +01:00
Clément Renault	58d039a70d	Introduce the FacetIter Iterator	2020-12-04 12:02:23 +01:00
Clément Renault	d8e25a0863	Order documents by the first custom criterion on basic searches	2020-12-04 12:02:23 +01:00
Clément Renault	e0cc7faea1	Use the facet ordered to the search	2020-12-04 12:02:23 +01:00
Clément Renault	61b383f422	Introduce the criteria update setting	2020-12-04 12:02:22 +01:00
Clément Renault	f8f33d35e0	Add the criteria list to the index	2020-12-02 11:21:26 +01:00
Kerollmops	57e8e5c965	Move the FacetCondition to its own module	2020-12-02 11:21:26 +01:00
Clément Renault	ecc8bc8910	Introduce the FieldId u8 alias type	2020-12-02 11:19:45 +01:00
Clément Renault	0a63e69e04	Merge pull request #45 from meilisearch/infos-export-documents Infos export documents	2020-12-02 10:50:54 +01:00
Clément Renault	16755b26e2	Make the export words FST export infos subcommand outputs to stdout	2020-12-02 10:43:22 +01:00
Kerollmops	85d51ab228	Introduce an infos subcommand to export documents from an index	2020-12-02 10:42:48 +01:00
Clément Renault	92f253adb2	Merge pull request #41 from meilisearch/update-store-delete-updates Allow users to abort pending updates	2020-12-01 14:56:00 +01:00
Clément Renault	222f2913c1	Simplify the processing_update UpdateStore method	2020-12-01 14:51:05 +01:00
Kerollmops	878b1873cd	Make sure to avoid removing the first pending update as it is processed	2020-12-01 14:51:05 +01:00
Clément Renault	96f64c629e	Move the UpdateStore out of the update module	2020-12-01 14:51:05 +01:00
Clément Renault	58a1f9081c	Allow users to abort pending updates, one by one or all at once	2020-12-01 14:51:05 +01:00
Clément Renault	e4c2abb1d9	Merge pull request #44 from meilisearch/clippy Fix some clippy warnings	2020-12-01 14:50:31 +01:00
Kerollmops	d0240bd9d0	Done a big clippy pass	2020-12-01 14:45:19 +01:00
Clément Renault	6e3f4e5e45	Merge pull request #43 from meilisearch/lowercase-facet-strings Lowercase the facet string value	2020-12-01 14:44:39 +01:00
Kerollmops	844a9022fb	Introduce the FacetStringOperator equal and not_equal constructors	2020-12-01 14:29:44 +01:00
Kerollmops	45877b3154	Lowercase the facet string value	2020-12-01 14:10:00 +01:00
Clément Renault	6120f6590b	Merge pull request #38 from meilisearch/facet-queries Introduce a facet filter system	2020-11-28 17:21:07 +01:00
Clément Renault	ba4ba685f9	Make the facet levels maps to previous level groups and don't split them	2020-11-28 12:43:43 +01:00
Clément Renault	276c87af68	Introduce more test to the FacetCondition struct	2020-11-23 16:43:57 +01:00
Clément Renault	a50f63840f	Return spanned pest error while parsing numbers in facet filters	2020-11-23 16:43:57 +01:00
Clément Renault	54d5cec582	Transform numbers into strings when faceted and necessary	2020-11-23 16:43:56 +01:00
Clément Renault	fc686aaca7	Use the De Morgan law to simplify the NOT operation	2020-11-23 16:43:56 +01:00
Clément Renault	7370ef8c5e	Add two simple test to the facet FacetCondition struct construction	2020-11-23 16:43:56 +01:00
Clément Renault	fc242f6e1f	Rewrite the FacetCondtion Debug impl in a defensive way	2020-11-23 16:43:56 +01:00
Clément Renault	a0adfb5e8e	Introduce a real pest parser and support every facet filter conditions	2020-11-23 16:43:55 +01:00
Clément Renault	c52d09d5b1	Support a basic version of the string facet query system	2020-11-23 16:43:55 +01:00
Clément Renault	498f0d8539	Output the documents count for each facet value in the infos subcommand	2020-11-23 16:43:55 +01:00
Clément Renault	278391d961	Move the facets related system into the new search module	2020-11-23 16:43:54 +01:00
Clément Renault	531bd6ddc7	Make the facet operator evaluation code generic	2020-11-23 16:43:54 +01:00
Clément Renault	d40dd3e4da	Reduce the amount of duplicated code to iterate over facet values	2020-11-23 16:43:54 +01:00
Clément Renault	07a0c82790	Bump heed to 0.10.4 to use be able to lazily decode roaring bitmaps	2020-11-23 16:43:53 +01:00
Clément Renault	59ca4b9fe4	Introduce a little bit of debug when deleting documents	2020-11-23 16:43:53 +01:00
Clément Renault	0694cc4916	Drastically speed up documents deletion updates	2020-11-23 16:43:53 +01:00
Clément Renault	38c76754ef	Make the facet level search system generic on f64 and i64	2020-11-23 16:43:52 +01:00
Clément Renault	9e2cbe3362	Improve the FacetLevelF64 serialization	2020-11-23 16:43:52 +01:00
Clément Renault	ced0c29c56	Simplify getting the biggest level of a facet field	2020-11-23 16:43:52 +01:00
Kerollmops	7d67c9e2e7	Improve the facet search algorithm performances	2020-11-23 16:43:52 +01:00
Clément Renault	67d4a1b3fc	Introduce a new update for the facet levels	2020-11-23 16:43:51 +01:00
Clément Renault	45e0feab4e	Speed up the facets stats infos subcommand	2020-11-23 16:43:51 +01:00
Kerollmops	7a6e6eb5e2	Introduce a facets stats infos subcommand	2020-11-23 16:43:51 +01:00
Clément Renault	9ec95679e1	Introduce a function to retrieve the facet level range docids	2020-11-23 16:43:50 +01:00
Clément Renault	57d253aeda	Improve the infos biggest-value subcommand to support facets	2020-11-23 16:43:50 +01:00
Clément Renault	fd8360deb1	Update the facet indexing facet test	2020-11-23 16:43:50 +01:00
Clément Renault	9b7e516a56	Fix the indexing process going back in time	2020-11-23 16:43:49 +01:00
Clément Renault	b255be93fa	Bump heed to 0.10.3	2020-11-23 16:43:49 +01:00
Clément Renault	218eb97241	Introduce an input field for the facet filters on the http-ui	2020-11-23 16:43:49 +01:00
Clément Renault	2341b99379	Support a basic facet based query system	2020-11-23 16:43:49 +01:00
Clément Renault	1d5795d134	Merge pull request #39 from meilisearch/speedup-documents-ids-merging Speedup documents ids merging	2020-11-22 19:32:24 +01:00
Clément Renault	05c95dfdc6	Introduce an infos subcommand that patches the external documents ids	2020-11-22 19:27:34 +01:00
Clément Renault	27f3ef5f7a	Use the new ExternalDocumentsIds struct in the engine	2020-11-22 19:27:34 +01:00
Clément Renault	fe82516f9f	Use the ExternalDocumentsIds in the Index struct	2020-11-22 19:27:34 +01:00
Clément Renault	415c0b86ba	Introduce the ExternalDocumentsIds struct	2020-11-22 19:27:33 +01:00
Clément Renault	eded5558b2	Rename the users ids documents ids into external documents ids	2020-11-22 17:17:47 +01:00
Clément Renault	f06355b0bb	Display the time it takes to merge user documents ids	2020-11-22 11:28:35 +01:00
Clément Renault	b0c5f59c07	Merge pull request #36 from meilisearch/index-facets Index facets values and support facet numbers	2020-11-14 14:32:05 +01:00
Clément Renault	e76558b0cc	Change the settings update system to reindex only one time	2020-11-14 11:17:49 +01:00
Clément Renault	f9cc12ae0f	Do not try to parse empty faceted strings	2020-11-13 18:35:47 +01:00
Clément Renault	23f9a22edc	Update the HTTP settings route to accept the faceted fields	2020-11-13 18:35:47 +01:00
Clément Renault	8e6efe4d87	Introduce an infos subcommand to display the facet values	2020-11-13 18:35:47 +01:00
Clément Renault	a18d9a1f87	Parse and store the faceted fields	2020-11-13 16:13:51 +01:00
Clément Renault	4e5e55c21a	Simplify the merge functions	2020-11-13 14:50:30 +01:00
Clément Renault	8ae9888959	Store the field id instead of the field name in the facets database	2020-11-13 14:50:30 +01:00
Clément Renault	cf9ddd293d	Simplify the the facet types	2020-11-13 11:46:48 +01:00
Clément Renault	466fb601d6	Faceted fields settings must specify the facet type	2020-11-13 11:46:48 +01:00
Clément Renault	ebe7087bff	Introduce the faceted fields setting	2020-11-11 17:08:18 +01:00
Clément Renault	72f18759ba	Introduce getters and setters for the facet fields ids facet types	2020-11-11 16:26:22 +01:00
Clément Renault	92ec908303	Introduce the facet field id values engine database	2020-11-11 16:06:33 +01:00
Clément Renault	e0058c1125	Introduce codecs for facet types (string, f64, u64, i64)	2020-11-11 15:48:24 +01:00
Clément Renault	b4951c058b	Merge pull request #35 from meilisearch/better-update-progress Better update progress	2020-11-11 13:19:32 +01:00
Clément Renault	a71a96894d	Use the new indexing progress events in the http server	2020-11-11 13:14:24 +01:00
Clément Renault	ea43080548	Make the indexing process send the new progress step events	2020-11-11 13:13:08 +01:00
Clément Renault	e78b96a657	Introduce a more detailed progress status enum	2020-11-11 12:31:59 +01:00
Clément Renault	8a4794fc51	Merge pull request #34 from meilisearch/speedup-indexing Write the words pairs proximities directly into LMDB to speedup indexing	2020-11-11 11:30:28 +01:00
Clément Renault	535f8088d7	Write the words pairs proximities directly into LMDB to speedup indexing	2020-11-11 11:25:31 +01:00
Clément Renault	fbe8ec1fe7	Merge pull request #33 from meilisearch/speedup-CI Avoid compiling benchmarks and speedup the CI	2020-11-11 11:20:26 +01:00
Clément Renault	a55453e634	Avoid compiling benchmarks and speedup the CI	2020-11-11 11:14:57 +01:00
Clément Renault	5a6b62e77c	Merge pull request #32 from meilisearch/http-get-one-document Introduce a route to get one document	2020-11-11 11:14:00 +01:00
Clément Renault	63fab07047	Introduce a route to retrieve a document with its id	2020-11-11 11:04:11 +01:00
Clément Renault	c00fc6f8bb	Merge pull request #31 from meilisearch/improve-update-process Improve update process	2020-11-09 17:45:19 +01:00
Clément Renault	0cfeee13ee	Reduce the number of documents limit when update progress are sent	2020-11-09 17:34:52 +01:00
Clément Renault	cf8a6a042e	Display a real progress bar when updates are processed	2020-11-09 17:33:36 +01:00
Clément Renault	45ae086974	Make sure pending updates are process when restarting the UpdateStore	2020-11-09 17:33:07 +01:00
Clément Renault	8ffdfa72e3	Merge pull request #28 from meilisearch/highlight-json-value Make the engine able to highlight any json type	2020-11-09 10:23:22 +01:00
Clément Renault	4fb138c42e	Make sure we index all kind of JSON types	2020-11-06 16:35:07 +01:00
Clément Renault	640c7d748a	Modify the highlight function to support any JSON type	2020-11-05 13:59:32 +01:00
Clément Renault	c94bc59d7e	Introduce a function to transform an obk into a JSON	2020-11-05 13:57:29 +01:00
Clément Renault	b220885f42	Fix the milli logo in the README	2020-11-05 11:43:47 +01:00
Clément Renault	1c2d36d8a3	Merge pull request #27 from meilisearch/split-http-ui Move the http server into its own sub-module	2020-11-05 11:36:04 +01:00
Clément Renault	0408c9d66a	Move the http server into its own sub-module	2020-11-05 11:16:39 +01:00
Clément Renault	749764f35b	Merge pull request #26 from meilisearch/searchable-attributes Introduce the searchable attributes	2020-11-04 09:40:03 +01:00
Clément Renault	a31db33e93	Introduce an optimization when the searchable attributes are ordered	2020-11-03 19:59:09 +01:00
Clément Renault	01c4f5abcd	Introduce the searchable attributes setting to the settings route	2020-11-03 19:35:55 +01:00
Clément Renault	63f65bac3e	Ignore the long running UpdateStore test	2020-11-03 19:12:00 +01:00
Clément Renault	a20c871ece	Add more tests to the Settings searchable attributes operation	2020-11-03 18:58:19 +01:00
Clément Renault	649fb6e401	Make sure that the indexing Store only index searchable fields	2020-11-03 18:58:19 +01:00
Clément Renault	e48630da72	Introduce the searchable parameter settings to the Settings update	2020-11-03 18:58:19 +01:00
Clément Renault	68d783145b	Introduce searchable fields methods on the index	2020-11-03 18:58:19 +01:00
Clément Renault	32486b5beb	Merge pull request #25 from meilisearch/update-ci Update the Github Actions settings	2020-11-03 18:53:04 +01:00
Clément Renault	a716ec61b9	Remove the fmt and clippy jobs	2020-11-03 18:52:45 +01:00
Clément Renault	c059924a8f	Remove the bors config at it does not work on private repositories	2020-11-03 18:25:49 +01:00
Clément Renault	3ef031b2fe	Update the CI to work on push and PRs	2020-11-03 18:25:12 +01:00
Clément Renault	58c07e7f8c	Merge pull request #23 from meilisearch/update-builder-thread-pool Allow library users to specify the rayon ThreadPool for UpdateBuilder	2020-11-02 19:11:50 +01:00
Clément Renault	7e120fc441	Allow library users to specify the rayon ThreadPool for UpdateBuilder	2020-11-02 19:11:22 +01:00
Clément Renault	87902de010	Merge pull request #22 from meilisearch/update-readme Update the README	2020-11-02 18:28:16 +01:00
Clément Renault	1718fe3d74	Update the README to be up to date with the recent updates	2020-11-02 18:07:24 +01:00
Clément Renault	82322ddab6	Merge pull request #21 from meilisearch/displayed-attributes Add the displayed attributes setting to an index	2020-11-02 15:50:29 +01:00
Clément Renault	3d1854ab95	Introduce an HTTP route to accept settings changes	2020-11-02 15:47:21 +01:00
Clément Renault	995d72b8c1	Introduce the Settings update operation	2020-11-02 15:31:20 +01:00
Clément Renault	0c612f08c7	Rename the indexing warp routes	2020-11-02 15:30:29 +01:00
Clément Renault	9b08f48dbd	Construct the documents based on the displayed fields or fields ids order	2020-11-02 13:01:32 +01:00
Clément Renault	303c3ce89e	Clean up the heed imports in the index module	2020-11-02 12:49:54 +01:00
Clément Renault	8f56753a2f	Introduce displayed fields methods on the index	2020-11-02 12:49:54 +01:00
Clément Renault	4fded5bd0e	Bump heed to be able to reference a RoTxn from multiple threads	2020-11-02 12:49:23 +01:00
Clément Renault	3abfe8aa22	Validate documents ids before accepting them	2020-11-01 20:55:21 +01:00
Clément Renault	0ccf4cf785	Simplify the IndexDocuments builder creation from the UpdateBuilder	2020-11-01 17:31:20 +01:00
Clément Renault	d8ff939409	Introduce bors to the project	2020-11-01 14:49:07 +01:00
Clément Renault	9047dc8163	Add a Github actions workflows	2020-11-01 14:47:44 +01:00
Clément Renault	600aa223c2	Fix a bug where generated docids were not saved when indexing JSON docs	2020-11-01 12:19:07 +01:00
Clément Renault	f0e63025b0	Update the Transform struct to support JSON stream updates	2020-11-01 12:19:06 +01:00
Kerollmops	082ad84914	Fix the benchmarks	2020-10-31 22:18:29 +01:00
Kerollmops	6d52c5b2f0	Introduce a parameter to disable the engine to autogenerate docids	2020-10-31 21:46:55 +01:00
Clément Renault	21b4d60101	Add replace/update csv/json from the HTTP server	2020-10-31 20:52:49 +01:00
Clément Renault	a4f8be7811	Support numbers and boolean when indexing JSON	2020-10-31 20:52:49 +01:00
Clément Renault	f0d028d3a4	Update the Transform struct to support JSON updates	2020-10-31 20:52:49 +01:00
Clément Renault	9d47ee52b4	Generate a uuid v4 based document id when missing	2020-10-31 15:11:06 +01:00
Clément Renault	ddbd336387	Introduce primary key methods on the index	2020-10-31 11:50:59 +01:00
Clément Renault	0d01e4854b	Add a test to check that merging works correctly with CSVs	2020-10-30 13:46:56 +01:00
Clément Renault	955302fd95	Introduce an HTTP route to clear the documents	2020-10-30 13:12:55 +01:00
Clément Renault	7cc1a358f5	Fix a documents indexing bug and add a test	2020-10-30 12:14:25 +01:00
Clément Renault	99da69c85f	Introduce the prepare_for_closing Index method	2020-10-30 11:46:14 +01:00
Clément Renault	222063b19d	Introduce the Index path method	2020-10-30 11:46:00 +01:00
Clément Renault	085d3b9d94	Update heed to 0.10.0	2020-10-30 11:42:00 +01:00
Clément Renault	a30206a665	Prefer using iterator put_current instead of a get put method	2020-10-30 11:13:45 +01:00
Clément Renault	e63fdf2b22	Move the heed env into the index itself to ease the usage of the library	2020-10-30 10:56:35 +01:00
Clément Renault	b5d52b6b45	Prefer using a smallstr instead of a real String to reduce allocations	2020-10-29 14:32:32 +01:00
Clément Renault	40993a0d25	Fix an indexing process bug, where documents were not written in order	2020-10-29 14:20:03 +01:00
Clément Renault	855a251489	Enable the clear documents optimization that wasn't working due to a bug	2020-10-29 13:52:48 +01:00
Clément Renault	1228c2948d	Add a comment about the ClearDocuments operation in the DeleteDocuments	2020-10-28 11:17:36 +01:00
Clément Renault	98fc24cbdf	Bump heed to fix a prefix iter bug	2020-10-28 10:55:21 +01:00
Kerollmops	d6338af766	Improve documents deletion by iterating over all the word pair positions	2020-10-27 18:50:09 +01:00
Clément Renault	3889d956d9	Introduce the UpdateBuilder and use it in the HTTP routes	2020-10-27 18:47:58 +01:00
Clément Renault	5c62fbb6a8	Move the IndexDocuments update into its own module	2020-10-26 12:21:13 +01:00
Clément Renault	8f76ec97c0	Move the DeleteDocuments update into its own module	2020-10-26 11:01:00 +01:00
Clément Renault	92ef1faa97	Move the ClearDocuments update into its own module	2020-10-26 10:58:17 +01:00
Clément Renault	1e1821f002	Introduce the merge_two_obkv function to merge documents on update	2020-10-26 10:55:07 +01:00
Clément Renault	60347a5483	Move the AvailableDocumentsIds iterator into the update module	2020-10-26 10:53:23 +01:00
Clément Renault	b14cca2ad9	Introduce the UpdateBuilder type along with some update operations	2020-10-25 18:32:01 +01:00
Clément Renault	adacc7977d	Make the Index return default values when value don't exist	2020-10-25 18:30:24 +01:00
Clément Renault	a7a4984175	Introduce the Transform type into the indexing system	2020-10-24 17:06:09 +02:00
Clément Renault	b44b04d25b	Serialize the CSV record values as JSON strings	2020-10-24 14:43:46 +02:00
Clément Renault	656a851830	Introduce the Transform struct transforming CSVs This allows us to: - transform a CSV, a JSON or a JSON lines data type into the same Grenad x Obkv streamable data type and creates the new FieldsIdsMap. - Extract all the documents user ids in advance to be able to delete the existing documents before re-indexing them. - Keep the last documents with the same user id avoiding duplicates in the same request.	2020-10-24 13:37:38 +02:00
Clément Renault	8d82e37ec0	Introduce the AvailableDocumentsIds iterator	2020-10-23 12:07:01 +02:00
Clément Renault	2a4cd81c86	Add documentation to the Index methods	2020-10-22 15:44:12 +02:00
Clément Renault	566a7c3039	Make the FieldsIdsMap serialization more stable by using a BTreeMap	2020-10-22 14:53:20 +02:00
Clément Renault	9133f38138	Introduce the FieldsIdsMap type	2020-10-22 12:56:35 +02:00
Clément Renault	802e925fd7	Switch to a JSON protocol for the front page	2020-10-21 18:26:29 +02:00
Clément Renault	5caf523fd9	Move the Index to its own module	2020-10-21 15:55:48 +02:00
Clément Renault	2210818114	Introduce the obkv heed codec	2020-10-21 15:51:48 +02:00
Clément Renault	f6eecb855e	Send a basic progressing status to the updates front page	2020-10-21 15:38:28 +02:00
Clément Renault	4eeeccb9cd	Change the UpdateStore to have different processed and pending meta types	2020-10-21 13:52:15 +02:00
Clément Renault	16ab3e02a9	Change the UpdateStore internal meta serializer	2020-10-21 13:42:49 +02:00
Clément Renault	f948a03be2	Optimise the merge functions to avoid allocations	2020-10-20 16:40:50 +02:00
Clément Renault	cde8478388	Replace the panic in the merge function by actual errors	2020-10-20 16:19:07 +02:00
Clément Renault	8ed8abb9df	Introduce an append-only indexing system	2020-10-20 15:00:58 +02:00
Clément Renault	a122d3d466	Export the indexing part into a module	2020-10-20 14:22:09 +02:00
Clément Renault	eb92e72e6c	Updates can send progress update status	2020-10-20 12:28:10 +02:00
Clément Renault	341046c96c	Remove the js map file from the filesize.js script	2020-10-20 12:20:42 +02:00
Clément Renault	3a934b7020	Split the update attributes on the updates front page	2020-10-20 12:19:48 +02:00
Clément Renault	03ca1ff634	Make the updates page interactive	2020-10-20 12:09:38 +02:00
Clément Renault	35c9a3c558	Brodacast the updates infos to every ws clients	2020-10-20 11:19:34 +02:00
Clément Renault	56c3a61d83	Introduce a new updates page	2020-10-19 19:57:15 +02:00
Clément Renault	871222aebd	Introduce some new routes to handle live indexing	2020-10-19 16:06:43 +02:00
Clément Renault	d3145be744	Rename the meta UpdateStore method	2020-10-19 14:00:00 +02:00
Clément Renault	8bfa43f9a7	Update the iter_metas UpdateStore method	2020-10-19 13:58:08 +02:00
Clément Renault	65e32fecb1	Move the binaries into one with subcommands	2020-10-19 13:44:17 +02:00
Clément Renault	ff389f1270	Update heed-types to 0.7.1	2020-10-19 11:52:59 +02:00
Clément Renault	5b4eda670b	Add two tests for the UpdateStore	2020-10-18 18:55:09 +02:00
Clément Renault	edb8c99fbe	Introduce a method to get the meta of an update on the UpdateStore	2020-10-18 17:19:04 +02:00
Clément Renault	eca49e3a03	Introduce a notification channel for the UpdateStore	2020-10-18 16:37:37 +02:00
Clément Renault	83c1db8763	Introduce the UpdateStore	2020-10-18 15:26:57 +02:00
Clément Renault	90d4c1d153	Simplify the words pair proximity computation	2020-10-15 16:18:43 +02:00
Clément Renault	9021b2dba6	Introduce the enable-chunk-fusing flag	2020-10-14 18:44:59 +02:00
Kerollmops	f980422c57	Move from oxidized-mtbl to grenad	2020-10-14 12:47:32 +02:00
Clément Renault	b342a86c15	Divide the max-memory parameter by the number of sorters in the store	2020-10-08 17:27:53 +02:00
Kerollmops	fb2c402ae1	Split the max-memory by the number of jobs	2020-10-07 14:23:22 +02:00
Kerollmops	38820bc75c	Improve and simplify the query tokenizer	2020-10-07 14:23:22 +02:00
Kerollmops	4e9bd1fef5	Bump oxidized-mtbl	2020-10-07 14:23:22 +02:00
Kerollmops	a00f5850ee	Add support for placeholder search for empty queries	2020-10-06 20:19:50 +02:00
Kerollmops	433d9bbc6e	Use CompressionType::from_str rather than a custom function	2020-10-06 13:50:34 +02:00
Kerollmops	4b819457c9	Enable the strucopt/clap warp help feature	2020-10-06 13:06:22 +02:00
Clément Renault	a2182e68a6	Rewrite the parallel merge indexing part	2020-10-05 20:54:06 +02:00
Kerollmops	e9e03259c1	Improve the mDFS performance and return the proximity	2020-10-05 18:13:56 +02:00
Kerollmops	bb15f16d8c	Merge other databases content while writing into LMDB at the same time	2020-10-05 16:35:10 +02:00
Clément Renault	9af946a306	Merging the main, word docids and words pairs proximity docids in parallel	2020-10-04 18:40:34 +02:00
Clément Renault	99705deb7d	Directly use a writer for the docid word positions	2020-10-04 18:17:53 +02:00
Clément Renault	67577a3760	It is an error to merge docid word positions	2020-10-04 17:31:12 +02:00
Clément Renault	ce8e56ee18	Rewrite the indexer to use one MTBL by database This allows us to avoid prefixing keys and appending into LMDB databases	2020-10-04 17:04:33 +02:00
Clément Renault	770f29fd05	Bump the oxidized-mtbl dependency	2020-10-04 17:04:33 +02:00
Clément Renault	acd2a63879	Introduce a simple FST based chinese word segmenter	2020-10-04 17:04:33 +02:00
Clément Renault	6cc6addc2f	Increase the CboRoaringBitmapCodec threshold	2020-10-02 17:06:17 +02:00
Clément Renault	e41a3822a6	Add a simple test for the CboRoaringBitmapCodec	2020-10-02 16:52:36 +02:00
Clément Renault	c4b0c57059	Reduce the default indexer max-memory parameter	2020-10-02 16:47:41 +02:00
Kerollmops	007e647462	Introduce the Mdfs Iterator that explore the proximity graph using a mana DFS	2020-10-02 16:46:07 +02:00
Kerollmops	d4e80407e5	Introduce the mana depth first search algorithm	2020-10-02 16:46:07 +02:00
Kerollmops	f6a8096720	Rename the quartile as percentiles 25th, 50th and 75th	2020-10-02 16:46:07 +02:00
Kerollmops	891e0188dd	Introduce the database-stats infos subcommand	2020-10-02 16:46:07 +02:00
Kerollmops	079742b4d3	Clean up the stats and size of database infos subcommands	2020-10-02 16:46:06 +02:00
Kerollmops	d0c73564b1	Use the CboRoaringBitmapCodec for the word pair proximity docids	2020-10-02 16:46:06 +02:00
Kerollmops	5a6a698e1d	Introduce the CboRoaringBitmapCodec	2020-10-02 16:46:06 +02:00
Kerollmops	4eda149ffa	Rename the BoRoaringBitmap codec	2020-10-02 16:46:06 +02:00
Clément Renault	ac84db2506	Move the words pairs proximities average into the stats infos subcommand	2020-10-02 16:46:06 +02:00
Kerollmops	30755e31e7	Introduce the words pairs proximities stats info subcommand	2020-10-02 16:46:06 +02:00
Clément Renault	bc35c9a598	Introduce the size_of_database infos subcommand	2020-10-02 16:46:05 +02:00
Kerollmops	c6b883289c	Remove the unused fetch_keywords function	2020-09-30 15:41:23 +02:00
Kerollmops	58237bd67f	Introduce the average-number-of-document-by-word-pair-proximity infos subcommand	2020-09-29 18:32:48 +02:00
Kerollmops	991be8950e	Rename the subcommand into average-number-of-positions-by-word-by-doc	2020-09-29 18:15:44 +02:00
Kerollmops	54370e228a	Search for documents with longer proximities until we find enough	2020-09-29 17:37:14 +02:00
Kerollmops	f277ea134f	Simplify some search function by reducing the number of parameters	2020-09-29 16:08:58 +02:00
Kerollmops	68f4af7d2e	Improve the display of the number of processed documents	2020-09-29 16:08:58 +02:00
Kerollmops	59a127d022	Improve the indexing process We now store the words pairs proximity in a cache and only compute the shortest proximity between pairs of words in a document.	2020-09-29 15:09:18 +02:00
Kerollmops	6ddb3e722c	Depth-first search cache the docids unions	2020-09-28 16:55:21 +02:00
Kerollmops	a3821a0b33	Introduce the depth_first_search path resolution function	2020-09-28 16:34:12 +02:00
Kerollmops	51c237f9d8	Fix the benchmarks compilation	2020-09-28 13:39:17 +02:00
Clément Renault	d8354f6f02	Fix the word_docids capacity limit detection	2020-09-27 11:52:05 +02:00
Clément Renault	25b2853b70	Move the words pairs proximities compute into the write document function	2020-09-23 15:02:40 +02:00
Clément Renault	ed05999f63	Replace the arc cache by a simple linked hash map	2020-09-23 14:50:52 +02:00
Clément Renault	4d22d80281	Display only the key on heed error	2020-09-23 14:13:51 +02:00
Clément Renault	5178b3d59d	Make the search system be aware of query words typos	2020-09-23 12:01:39 +02:00
Clément Renault	b597a92487	Add a default max-memory value to the indexer	2020-09-23 12:00:36 +02:00
Clément Renault	1f6e00878d	Use the words pair proximities in the search algorithm	2020-09-22 18:47:55 +02:00
Clément Renault	31224a8425	Index the word pair proximities for both orders of the pair	2020-09-22 14:49:22 +02:00
Clément Renault	a58ae5eb2a	Introduce the word-pair-proximities-docids infos subcommand	2020-09-22 14:04:34 +02:00
Clément Renault	d6fa9c0414	Index the intra documents word pair proximities	2020-09-22 14:04:33 +02:00
Clément Renault	7b67ae6972	Introduce the StrStrU8 heed codec	2020-09-22 12:44:17 +02:00
Clément Renault	e34437b2d7	Move the proximity function to a module	2020-09-22 10:54:59 +02:00
Clément Renault	15208c7d3d	Simplify the indexer record loop	2020-09-22 10:33:30 +02:00
Clément Renault	e5adfaade0	Replace the token filter by a filter mapper	2020-09-22 10:24:31 +02:00
Clément Renault	d21c80b865	Apply the chunk compression parameters on all the MTBL writers	2020-09-21 18:30:54 +02:00
Clément Renault	944df52e2a	Simplify the indexer main loop	2020-09-21 14:59:48 +02:00
Kerollmops	3ded98e5fa	Bump the roaring version that fix a deserialization bug	2020-09-10 22:37:51 +02:00
Kerollmops	d5e5baa20f	Bump the oxidized-mtbl dependency	2020-09-10 13:29:12 +02:00
Kerollmops	0fb086f241	Use the crates.io raoring library	2020-09-08 15:16:04 +02:00
Kerollmops	aed0704404	Remove the temporary optimisation	2020-09-08 14:48:33 +02:00
Kerollmops	072382fa61	Sort the word docids to make intersections much faster	2020-09-07 22:38:49 +02:00
Kerollmops	ad11c5fb3f	Introduce the words-docids command for the infos binary	2020-09-07 22:36:35 +02:00
Kerollmops	5664c37539	Introduce an heed codec that reduce the size of small amount of serialized integers	2020-09-07 20:06:23 +02:00
Kerollmops	3e2250423c	Introduce the average-number-of-positions infos subcommand	2020-09-07 15:26:42 +02:00
Kerollmops	ea605b499c	Introduce two new infos subcommands	2020-09-07 14:56:48 +02:00
Clément Renault	bb1ab428db	Use another function to define the proximity	2020-09-06 17:55:07 +02:00
Clément Renault	f928b91e9d	Specify the exact rev for the near-proximity dep	2020-09-06 17:21:38 +02:00
Clément Renault	dec460ce52	Fix the infos binary and add commands	2020-09-06 17:14:20 +02:00
Clément Renault	daa3673c1c	Invert the word docid positions key order	2020-09-06 10:30:53 +02:00
Clément Renault	c2405bcae2	Prefer using the word_docids db to create the words-fst	2020-09-06 10:23:56 +02:00
Kerollmops	4ca9472e02	Fix the minimum proximity len	2020-09-06 10:19:34 +02:00
Clément Renault	1c504471d3	Introduce the plane-sweep algorithm	2020-09-05 18:25:27 +02:00
Clément Renault	dc88a86259	Store the word positions under the documents	2020-09-05 18:03:06 +02:00
Kerollmops	580ed1119a	Make the engine to return csv string records as documents and headers	2020-08-31 19:02:00 +02:00
Clément Renault	bad0663138	Come back to the old tokenizer	2020-08-31 13:34:38 +02:00
Kerollmops	220ba0785c	Make the front-end to throttle the request by 100ms	2020-08-31 13:34:35 +02:00
Clément Renault	4afc4d0751	Use the groups of four positions to speed up disjunctions tests	2020-08-30 16:25:11 +02:00
Clément Renault	605f75b56f	Add the words grouped by four positions in the infos binary	2020-08-29 18:23:33 +02:00
Clément Renault	ad5cafbfed	Introduce a database to store docids in groups of four positions	2020-08-29 17:42:55 +02:00
Clément Renault	3db517548d	Move the documents back into the LMDB database	2020-08-29 15:14:04 +02:00
Clément Renault	816db7a0aa	Improve the RoaringBitmap codec to reserve enough vector space	2020-08-29 11:21:30 +02:00
Clément Renault	3fe497e129	Improve the Mtbl heed codec to only encode MTBL databases	2020-08-29 11:20:39 +02:00
Clément Renault	21aafd603c	Make sure the first document is associated to the document id 0	2020-08-29 10:56:40 +02:00
Clément Renault	0a44ff86ab	Put the documents MTBL back into LMDB We makes sure to write the documents into a file before memory mapping it and putting it into LMDB, this way we avoid moving it to RAM	2020-08-28 15:43:24 +02:00
Clément Renault	d784d87880	Remove the prefix LMDB databases	2020-08-28 14:41:43 +02:00
Clément Renault	7cde312f14	Introduce the StrBEU32Codec heed codec	2020-08-28 14:16:37 +02:00
Clément Renault	34db376ae5	Rename the RoaringBitmapCodec module	2020-08-28 13:31:16 +02:00
Kerollmops	38ddc71b83	Simplify the search algorithm	2020-08-26 15:16:41 +02:00
Kerollmops	ba2eb0d7ad	Take the words-fst into account when retrieving the biggests values	2020-08-26 14:36:22 +02:00
Clément Renault	32da07ccee	Introduce the word-positions-doc-ids and words-positions infos commands	2020-08-23 10:52:47 +02:00
Clément Renault	d19f394630	Make the indexer support gzipped CSV as input	2020-08-21 18:10:24 +02:00
Clément Renault	ff479c865d	Replace pipe by ringtail to improve stdin read performances	2020-08-21 17:45:52 +02:00
Clément Renault	ada30c2789	Introducing more arguments to specify the different compression algorithms	2020-08-21 16:41:26 +02:00
Clément Renault	02335ee72d	Introduce the biggest-value-sizes command on the infos binary	2020-08-21 14:44:42 +02:00
Clément Renault	1e3e756c19	Introduce the words-frequencies command on the infos binary	2020-08-21 14:44:42 +02:00
Kerollmops	6a230fe803	Move the contains_documents logic to a function	2020-08-21 14:44:42 +02:00
Kerollmops	e55a569629	Compress much more the documents database	2020-08-21 14:44:42 +02:00
Kerollmops	962bad3cea	Introduce an infos binary to fetch stats	2020-08-17 19:41:49 +02:00
Clément Renault	8806fcd545	Introduce a better query and document lexer	2020-08-16 14:36:54 +02:00
Clément Renault	1e358e3ae8	Introduce the AstarBagIter that iterates through best paths	2020-08-15 16:24:06 +02:00
Clément Renault	7dc594ba4d	Introduce the Search builder struct	2020-08-13 14:27:51 +02:00
Clément Renault	bfb46cbfbe	Introduce the Crtierion enum	2020-08-12 10:43:02 +02:00
Clément Renault	6d04a285dc	Retrieve and display the distances of the words found	2020-08-11 15:18:02 +02:00
Clément Renault	1bd37d213a	Lowercase quoted words	2020-08-10 14:49:09 +02:00
Clément Renault	883a8109c8	Show both database and documents database sizes	2020-08-10 14:37:18 +02:00
Clément Renault	a4e0f3f724	Remove the useless TransitiveArc from the serve binary	2020-08-10 14:06:27 +02:00
Clément Renault	edc06a97d6	Remove the useless stats binary	2020-08-10 13:55:02 +02:00
Clément Renault	ae77fe5a69	Introduce an option to specify the maximum database size	2020-08-10 13:53:53 +02:00
Clément Renault	394844062f	Move the documents MTBL database inside the Index	2020-08-10 13:47:19 +02:00
Clément Renault	ecd2b2f217	Make the final merge done in parallel	2020-08-07 15:44:04 +02:00
Clément Renault	91282c8b6a	Move the documents into another file	2020-08-07 13:11:31 +02:00
Clément Renault	fae694a102	Put the documents into an MTBL database	2020-08-07 12:14:40 +02:00
Clément Renault	d5a356902a	Update oxidized-mtbl	2020-08-07 12:14:03 +02:00
Clément Renault	405a71d3a4	Accept csv from stdin	2020-08-06 13:38:21 +02:00
Clément Renault	d3b1096510	Compute the word attribute postings lists on each threads	2020-08-06 11:50:27 +02:00
Clément Renault	8d734941af	Clean up some lines	2020-08-06 10:20:26 +02:00
Clément Renault	a4e3c7c37c	Force the Papa parse delimiter	2020-08-05 14:11:46 +02:00
Clément Renault	6508d497ce	Replace the regex highlighting by a simple algorithm	2020-08-05 13:52:27 +02:00
Clément Renault	4873abe145	Introduce option flags to toggle the indexing engine	2020-08-05 12:10:41 +02:00
Clément Renault	bd4b18541c	Introduce a new indexer which uses an MTBL sorter	2020-08-04 15:44:37 +02:00
Clément Renault	3f21760d56	Update README.md	2020-08-04 15:40:37 +02:00
Clément Renault	bc3a0ac6a3	Display the milli logo and update the description	2020-08-04 15:40:02 +02:00
Kerollmops	d7d8f38fb7	Update bulma to spread the logo more	2020-07-16 23:45:02 +02:00
Kerollmops	ee305c9284	Replace the title by the milli logo	2020-07-15 23:55:28 +02:00
Kerollmops	9ade00e27b	Highlight all the matching words	2020-07-14 11:53:21 +02:00
Kerollmops	085c376655	Use the regex crate to highlight "hello"	2020-07-14 11:28:40 +02:00
Kerollmops	dd385ad05b	Customize the mark tag css	2020-07-14 11:03:21 +02:00
Kerollmops	aa92311d4e	Add a dark theme to the dashboard	2020-07-13 23:51:41 +02:00
Kerollmops	3d144e62c4	Search for best proximities in multiple attributes	2020-07-13 19:06:56 +02:00
Kerollmops	576dd011a1	Compute the candidates but not by attribute	2020-07-13 18:16:05 +02:00
Kerollmops	6b14b20369	Introduce a method to retrieve the number of attributes of the documents	2020-07-13 17:50:16 +02:00
Kerollmops	54afec58a3	Add a fade in out animation when the server process	2020-07-12 11:34:48 +02:00
Kerollmops	92c2b1dd2d	Refine the help message of the binaries	2020-07-12 11:06:45 +02:00
Kerollmops	f757df5dfd	Introduce the stderr logger to the project	2020-07-12 11:04:35 +02:00
Kerollmops	12358476da	Use the log crate instead of stderr	2020-07-12 10:55:09 +02:00
Kerollmops	2c62eeea3c	Rename the project milli	2020-07-12 00:16:41 +02:00
Kerollmops	d31da26a51	Avoid cloning RoraringBitmaps when unecessary	2020-07-11 23:51:32 +02:00
Kerollmops	b8a1fc0126	Clean up the CSS style custom bulma rules	2020-07-11 14:51:59 +02:00
Kerollmops	f6eae91c7d	Pretty print the new dashboard numbers	2020-07-11 14:17:37 +02:00
Kerollmops	d44428fa90	Display more informations on the dashboard	2020-07-11 11:51:56 +02:00
Kerollmops	11c7fef80a	Implement a memory dumper It moves the in memory HashMaps used when indexing to a disk based MTBL file	2020-07-07 16:48:49 +02:00
Kerollmops	b12bfcb03b	Reduce the deepness of the word position document ids This helps reduce the number of allocations.	2020-07-07 12:30:05 +02:00
Kerollmops	7178b6c2c4	First basic version using MTBL again	2020-07-07 11:32:33 +02:00
Kerollmops	45d0d7c3d4	Clean up the README	2020-07-06 17:38:22 +02:00
Kerollmops	adb1038b26	Add a `jobs` parameter to set the number of threads the indexer uses	2020-07-06 12:17:17 +02:00
Kerollmops	2a3b03138b	Use heed 0.8.1 with the RwIter append method	2020-07-05 19:50:28 +02:00
Kerollmops	ec1023e790	Intersect document ids by inverse popularity of the words This reduces the worst request we had which took 56s to now took 3s ("the best of the do").	2020-07-05 19:33:51 +02:00
Kerollmops	cd7e64b2b3	Allow users to set the arc cache size when indexing	2020-07-04 18:12:41 +02:00
Kerollmops	ac8353a64f	Merge pre-computed word attribute documents ids	2020-07-04 17:02:27 +02:00
Kerollmops	fea7cac206	Display the time it took to compute the word attribute documents ids	2020-07-04 15:18:38 +02:00
Kerollmops	46ced5c828	Introduce the RwIter append heed API	2020-07-04 12:34:10 +02:00
Kerollmops	7e7440c431	Finalize the LMDB indexing design	2020-07-01 22:45:43 +02:00
Kerollmops	2ae3f40971	Make the indexer ignore certain words This is a preparation for making the indexing fully parallel by making the indexer only be aware of certain words for each threads to avoid postings lists conflicts for each words	2020-07-01 17:49:46 +02:00
Kerollmops	a3ac2623d5	Introduce multiple functions to clean up the code	2020-07-01 17:24:55 +02:00
Kerollmops	ac5cc7ddad	Introduce an Iterator yielding owned entries for the LruCache	2020-07-01 17:21:52 +02:00
Kerollmops	014a25697d	Use only one ARC cache based on the words	2020-07-01 12:03:18 +02:00
Kerollmops	fc4013a43f	Fix the ARC cache	2020-07-01 10:35:07 +02:00
Kerollmops	2fcae719ad	Use another LRU impl which uses hashbrown	2020-06-29 22:26:06 +02:00
Kerollmops	f98b615bf3	Replace the LRU by an Arc cache	2020-06-29 20:48:57 +02:00
Kerollmops	07abebfc46	Introduce a (too big) LRU cache	2020-06-29 18:15:03 +02:00
Kerollmops	5f0088594b	Index by writing directly into LMDB	2020-06-29 13:54:47 +02:00
Kerollmops	8453828a65	Update the README	2020-06-28 12:40:08 +02:00
Kerollmops	63cbeca64e	Skip all derived words when too short	2020-06-28 12:13:12 +02:00
Kerollmops	736f0f7560	Use the proximity instead of the attributes when searching for <= 7 proximities	2020-06-28 12:13:12 +02:00
Kerollmops	fe3be8f18a	Replace the HashMap by a Vec for attributes documents ids	2020-06-28 12:13:12 +02:00
Kerollmops	6a2834f2b0	Add a `jobs` parameter to set the number of threads the indexer uses	2020-06-28 12:13:10 +02:00
Kerollmops	7e16afbdce	Ignore documents which are not part of the candidates when exploring with A*	2020-06-24 15:06:45 +02:00
Kerollmops	1c7a9a4132	Remove the found documents from the candidates list	2020-06-24 15:00:26 +02:00
Kerollmops	50169b9798	Compute the full list of ids we are willing to find by attribute	2020-06-24 14:48:04 +02:00
Kerollmops	374ec6773f	Introduce a database to store all docids for a word and attribute	2020-06-22 19:24:20 +02:00
Kerollmops	a044cb6cc8	Clean up the warnings for prefix postings	2020-06-22 18:10:31 +02:00
Kerollmops	ba3e805981	Document the Index types and the internal LMDB databases	2020-06-22 18:09:22 +02:00
Kerollmops	2f0e1afd16	Introduce the roaring bitmap heed codec	2020-06-22 17:56:07 +02:00
Kerollmops	8148210860	Use the cache when retrieving the documents at the end	2020-06-21 12:25:19 +02:00
Kerollmops	1628a31efa	Cache the unions of the derived words positions	2020-06-20 15:38:10 +02:00
Kerollmops	115e0142d9	Add a feature flags to enable the export of stats	2020-06-20 13:25:42 +02:00
Kerollmops	beb49b24f6	Skip looking at connections for proximity 0	2020-06-20 13:19:03 +02:00
Kerollmops	c84012d655	Accept queries from standard input when not given as argument	2020-06-20 12:01:15 +02:00
Kerollmops	d6705d5529	Introduce the criterion dependency to bench the engine	2020-06-19 18:32:25 +02:00
Kerollmops	55a8941922	Optimize things	2020-06-19 17:48:17 +02:00
Kerollmops	a3ca80d20d	Ignore every proximities bigger or equal to 8	2020-06-18 15:42:46 +02:00
Kerollmops	3577de04b8	Reduce the number of KV lookups to the sucessfulls only	2020-06-16 12:58:29 +02:00
Kerollmops	e974e6b3c9	Acquire search intersections metrics	2020-06-16 12:10:23 +02:00
Kerollmops	8db16ff306	Add a cache to the contains_documents success function	2020-06-14 13:39:39 +02:00
Kerollmops	a8cda248b4	Introduce a customized A* algorithm. This custom algo lazily compute the intersections between words, to avoid too much set operations and database reads	2020-06-14 12:51:57 +02:00
Kerollmops	69285b22d3	Check that an edges combination contains results	2020-06-13 11:16:02 +02:00
Kerollmops	b9cc6c10af	Introduce a function to ignore useless paths	2020-06-13 00:17:43 +02:00
Kerollmops	d02c5cb023	Fix node skipping by computing the accumulated proximity	2020-06-12 14:08:46 +02:00
Kerollmops	37a48489da	Reworked the best proximity algo a little bit	2020-06-12 12:53:08 +02:00
Kerollmops	302866ad73	Make the algo don't work with an astar	2020-06-11 17:43:06 +02:00
Kerollmops	0a83a86e65	Fix multiple bugs	2020-06-11 11:55:03 +02:00
Kerollmops	4e86ecf807	Retrieve the words before the intersect loops	2020-06-10 22:05:01 +02:00
Kerollmops	6ca3579cc0	Add more time debug measurements	2020-06-10 21:35:01 +02:00
Kerollmops	66a4b26811	Introduce a proximity based documents retriever	2020-06-10 16:54:28 +02:00
Kerollmops	78f27c0465	squash-me: Remove debugs	2020-06-10 16:29:46 +02:00
Kerollmops	3ad883d7c7	squash-me: Make the dijkstra work even with different attributes	2020-06-10 16:27:02 +02:00
Kerollmops	fecd8ca54a	squash-me: It works! we must remove the debug after having added more tests	2020-06-10 14:20:35 +02:00
Kerollmops	13977d9338	squash-me	2020-06-09 23:06:59 +02:00
Kerollmops	5d5b827f1a	Squash-me	2020-06-09 17:32:25 +02:00
Kerollmops	2a6d6a7f69	Introduce a first draft of the best_proximity algorithm	2020-06-09 10:11:43 +02:00
Kerollmops	dfdaceb410	Introduce a first basic working positions-based engine	2020-06-05 20:13:19 +02:00
Kerollmops	f51a63e4ef	Store documents ids under attribute ids	2020-06-05 16:32:14 +02:00
Kerollmops	ce86a43779	Make the query tokenizer a real Iterator	2020-06-05 09:49:28 +02:00
Kerollmops	f55f4cb02a	Not fetch the cached prefix postings when prefix is disabled	2020-06-04 21:22:45 +02:00
Kerollmops	06bf03f075	Add an help message on the front page aaa	2020-06-04 21:22:45 +02:00
Kerollmops	eefc6d7c44	Add support for quoted query phrases	2020-06-04 20:25:51 +02:00
Kerollmops	1f7035f18f	Just do a little clean-up	2020-06-04 19:13:28 +02:00
Kerollmops	71dc6a3828	Disable prefix search when query is ended by a whitespace	2020-06-04 18:37:20 +02:00
Kerollmops	5d1c625b74	Change the page index texts	2020-06-04 18:20:57 +02:00
Kerollmops	c42d3c19e2	Merge the whole list of generated MTBL in one go	2020-06-04 17:38:43 +02:00
Kerollmops	3a23dc242e	More efficiently merge MTBLs, more than two at a time	2020-06-04 16:17:24 +02:00
Kerollmops	1df1f88fe1	Directly write to LMDB without intermediate final MTBL	2020-06-01 21:30:39 +02:00
Kerollmops	2174042994	Merge only 3 MTBL at the same time	2020-06-01 19:49:58 +02:00
Kerollmops	5cc81a0179	Merge many MTBL into one a the same time	2020-06-01 18:39:58 +02:00
Kerollmops	6a047519f6	Do a merge two by two	2020-06-01 18:27:26 +02:00
Kerollmops	5404776f7a	Add a little bit more debug	2020-06-01 17:52:43 +02:00
Kerollmops	dff68a339a	Use OnceCell to cache levenshtein builders	2020-05-31 19:27:11 +02:00
Kerollmops	dde3e01a59	Introduce prefix postings ids for better perfs	2020-05-31 18:20:49 +02:00
Kerollmops	a26553c90a	Reintroduce a simple HTTP server	2020-05-31 17:48:13 +02:00
Kerollmops	2a10b2275e	Support prefix typo tolerant search	2020-05-31 17:18:13 +02:00
Kerollmops	ba9527abc0	Support typos with a levenshtein automata	2020-05-31 17:01:11 +02:00
Kerollmops	6c726df9b9	Support multiple space seperated words	2020-05-31 16:09:34 +02:00
Kerollmops	24587148fd	Introduce MTBL parallel merging before LMDB writing	2020-05-31 14:22:57 +02:00
Kerollmops	6762c2d08f	Clean up a little bit	2020-05-31 14:22:57 +02:00
Kerollmops	3a998cf39c	Far better usage of rayon to fold indexed data	2020-05-31 14:22:57 +02:00
Kerollmops	1237306ca8	Introduce a thread that write to heed	2020-05-31 14:22:57 +02:00
Kerollmops	3668627e03	Use zerocopy without bitpacking as a first step	2020-05-31 14:22:07 +02:00
Kerollmops	a81f201fad	Inroduce the use of RocksDB instead of sled (RAM)	2020-05-31 14:22:06 +02:00
Kerollmops	91ba938953	Initial commit	2020-05-31 14:22:06 +02:00
Clément Renault	4573f00a0d	Initial commit	2020-05-31 14:21:56 +02:00

1143 changed files with 142263 additions and 22103 deletions

									
										2

.cargo/config.toml
									
										Normal file
									
												View File
												
				@@ -0,0 +1,2 @@

				[alias]

				xtask = "run --release --package xtask --"

									
										3

.github/ISSUE_TEMPLATE/bug_report.md
									
										vendored
									
												View File
												
				@@ -23,7 +23,8 @@ A clear and concise description of what you expected to happen.

				**Screenshots**

				If applicable, add screenshots to help explain your problem.

				**Meilisearch version:** [e.g. v0.20.0]

				**Meilisearch version:**

				[e.g. v0.20.0]

				**Additional context**

				Additional information that may be relevant to the issue.

									
										11

.github/ISSUE_TEMPLATE/config.yml
									
										vendored
									
												View File
												
				@@ -1,10 +1,13 @@

				contact_links:

				  - name: Feature request & feedback

				  - name: Support questions & other

				    url: https://github.com/meilisearch/meilisearch/discussions/new

				    about: For any other question, open a discussion in this repository

				  - name: Language support request & feedback

				    url: https://github.com/meilisearch/product/discussions/categories/feedback-feature-proposal?discussions_q=label%3Aproduct%3Acore%3Atokenizer+category%3A%22Feedback+%26+Feature+Proposal%22

				    about:  The requests and feedback regarding Language support are not managed in this repository. Please upvote the related discussion in our dedicated product repository or open a new one if it doesn't exist.

				  - name: Any other feature request & feedback

				    url: https://github.com/meilisearch/product/discussions/categories/feedback-feature-proposal

				    about: The feature requests and feedback regarding the already existing features are not managed in this repository. Please open a discussion in our dedicated product repository

				  - name: Documentation issue

				    url: https://github.com/meilisearch/documentation/issues/new

				    about: For documentation issues, open an issue or a PR in the documentation repository

				  - name: Support questions & other

				    url: https://github.com/meilisearch/meilisearch/discussions/new

				    about: For any other question, open a discussion in this repository

									
										49

.github/ISSUE_TEMPLATE/sprint_issue.md
									
										vendored
									
										Normal file
									
												View File
												
				@@ -0,0 +1,49 @@

				---

				name: New sprint issue

				about: ⚠️ Should only be used by the engine team ⚠️

				title: ''

				labels: ''

				assignees: ''

				---

				Related product team resources: [PRD]() (_internal only_)

				Related product discussion:

				Related spec: WIP

				## Motivation

				<!---Copy/paste the information in PRD or briefly detail the product motivation. Ask product team if any hesitation.-->

				## Usage

				<!---Link to the public part of the PRD, or to the related product discussion for experimental features-->

				## TODO

				<!---Feel free to adapt this list with more technical/product steps-->

				- [ ] Release a prototype

				- [ ] If prototype validated, merge changes into `main`

				- [ ] Update the spec

				### Reminders when modifying the Setting API

				<!--- Special steps to remind when adding a new index setting -->

				- [ ] Ensure the new setting route is at least tested by the [`test_setting_routes` macro](https://github.com/meilisearch/meilisearch/blob/5204c0b60b384cbc79621b6b2176fca086069e8e/meilisearch/tests/settings/get_settings.rs#L276)

				- [ ] Ensure Analytics are fully implemented

				  - [ ] `/settings/my-new-setting` configurated in the [`make_setting_routes` macro](https://github.com/meilisearch/meilisearch/blob/5204c0b60b384cbc79621b6b2176fca086069e8e/meilisearch/src/routes/indexes/settings.rs#L141-L165)

				  - [ ] global `/settings` route configurated in the [`update_all` function](https://github.com/meilisearch/meilisearch/blob/5204c0b60b384cbc79621b6b2176fca086069e8e/meilisearch/src/routes/indexes/settings.rs#L655-L751)

				- [ ] Ensure the dump serializing is consistent with the `/settings` route serializing, e.g., enums case can be different (`camelCase` in route and `PascalCase` in the dump)

				#### Special cases when adding a setting for an experimental feature

				- [ ] ⚠️ API stability: The setting does not appear on the main settings route when the feature has never been enabled (e.g. mark it `Unset` when returned from the index in this situation. See [an example](https://github.com/meilisearch/meilisearch/blob/7a89abd2a025606a42f8b219e539117eb2eb029f/meilisearch-types/src/settings.rs#L608))

				- [ ] The setting cannot be set when the feature is disabled, either by the main settings route or the subroute (see [`validate_settings` function](https://github.com/meilisearch/meilisearch/blob/7a89abd2a025606a42f8b219e539117eb2eb029f/meilisearch/src/routes/indexes/settings.rs#L811))

				- [ ] If possible, the setting is reset when the feature is disabled (hard if it requires reindexing)

				## Impacted teams

				<!---Ping the related teams. Ask for the engine manager if any hesitation-->

				<!---@meilisearch/docs-team when there is any API change, e.g. settings addition-->

									
										12

.github/dependabot.yml
									
										vendored
									
										Normal file
									
												View File
												
				@@ -0,0 +1,12 @@

				# Set update schedule for GitHub Actions only

				version: 2

				updates:

				  - package-ecosystem: "github-actions"

				    directory: "/"

				    schedule:

				      interval: "monthly"

				    labels:

				      - 'skip changelog'

				      - 'dependencies'

				    rebase-strategy: disabled

									
										132

.github/is-latest-release.sh
									
										vendored
									
												View File
											
				@@ -1,132 +0,0 @@

				#!/bin/sh

				# Checks if the current tag should be the latest (in terms of semver and not of release date).

				# Ex: previous tag -> v0.10.1

				#     new tag -> v0.8.12

				#     The new tag should not be the latest

				#     So it returns "false", the CI should not run for the release v0.8.2

				# Used in GHA in publish-docker-latest.yml

				# Returns "true" or "false" (as a string) to be used in the `if` in GHA

				# GLOBAL

				GREP_SEMVER_REGEXP='v\([0-9]*\)[.]\([0-9]*\)[.]\([0-9]*\)$' # i.e. v[number].[number].[number]

				# FUNCTIONS

				# semverParseInto and semverLT from https://github.com/cloudflare/semver_bash/blob/master/semver.sh

				# usage: semverParseInto version major minor patch special

				# version: the string version

				# major, minor, patch, special: will be assigned by the function

				semverParseInto() {

				    local RE='[^0-9]*\([0-9]*\)[.]\([0-9]*\)[.]\([0-9]*\)\([0-9A-Za-z-]*\)'

				    #MAJOR

				    eval $2=`echo $1 | sed -e "s#$RE#\1#"`

				    #MINOR

				    eval $3=`echo $1 | sed -e "s#$RE#\2#"`

				    #MINOR

				    eval $4=`echo $1 | sed -e "s#$RE#\3#"`

				    #SPECIAL

				    eval $5=`echo $1 | sed -e "s#$RE#\4#"`

				}

				# usage: semverLT version1 version2

				semverLT() {

				    local MAJOR_A=0

				    local MINOR_A=0

				    local PATCH_A=0

				    local SPECIAL_A=0

				    local MAJOR_B=0

				    local MINOR_B=0

				    local PATCH_B=0

				    local SPECIAL_B=0

				    semverParseInto $1 MAJOR_A MINOR_A PATCH_A SPECIAL_A

				    semverParseInto $2 MAJOR_B MINOR_B PATCH_B SPECIAL_B

				    if [ $MAJOR_A -lt $MAJOR_B ]; then

				        return 0

				    fi

				    if [ $MAJOR_A -le $MAJOR_B ] && [ $MINOR_A -lt $MINOR_B ]; then

				        return 0

				    fi

				    if [ $MAJOR_A -le $MAJOR_B ] && [ $MINOR_A -le $MINOR_B ] && [ $PATCH_A -lt $PATCH_B ]; then

				        return 0

				    fi

				    if [ "_$SPECIAL_A"  == "_" ] && [ "_$SPECIAL_B"  == "_" ] ; then

				        return 1

				    fi

				    if [ "_$SPECIAL_A"  == "_" ] && [ "_$SPECIAL_B"  != "_" ] ; then

				        return 1

				    fi

				    if [ "_$SPECIAL_A"  != "_" ] && [ "_$SPECIAL_B"  == "_" ] ; then

				        return 0

				    fi

				    if [ "_$SPECIAL_A" < "_$SPECIAL_B" ]; then

				        return 0

				    fi

				    return 1

				}

				# Returns the tag of the latest stable release (in terms of semver and not of release date)

				get_latest() {

				    temp_file='temp_file' # temp_file needed because the grep would start before the download is over

				    curl -s 'https://api.github.com/repos/meilisearch/meilisearch/releases' > "$temp_file"

				    releases=$(cat "$temp_file" | \

				        grep -E "tag_name|draft|prerelease" \

				        | tr -d ',"' | cut -d ':' -f2 | tr -d ' ')

				        # Returns a list of [tag_name draft_boolean prerelease_boolean ...]

				        # Ex: v0.10.1 false false v0.9.1-rc.1 false true v0.9.0 false false...

				    i=0

				    latest=""

				    current_tag=""

				    for release_info in $releases; do

				        if [ $i -eq 0 ]; then # Cheking tag_name

				            if echo "$release_info" | grep -q "$GREP_SEMVER_REGEXP"; then # If it's not an alpha or beta release

				                current_tag=$release_info

				            else

				                current_tag=""

				            fi

				            i=1

				        elif [ $i -eq 1 ]; then # Checking draft boolean

				            if [ "$release_info" = "true" ]; then

				                current_tag=""

				            fi

				            i=2

				        elif [ $i -eq 2 ]; then # Checking prerelease boolean

				            if [ "$release_info" = "true" ]; then

				                current_tag=""

				            fi

				            i=0

				            if [ "$current_tag" != "" ]; then # If the current_tag is valid

				                if [ "$latest" = "" ]; then # If there is no latest yet

				                    latest="$current_tag"

				                else

				                    semverLT $current_tag $latest # Comparing latest and the current tag

				                    if [ $? -eq 1 ]; then

				                        latest="$current_tag"

				                    fi

				                fi

				            fi

				        fi

				    done

				    rm -f "$temp_file"

				    echo $latest

				}

				# MAIN

				current_tag="$(echo $GITHUB_REF | tr -d 'refs/tags/')"

				latest="$(get_latest)"

				if [ "$current_tag" != "$latest" ]; then

				    # The current release tag is not the latest

				    echo "false"

				else

				    # The current release tag is the latest

				    echo "true"

				fi

									
										41

.github/scripts/check-release.sh
									
										vendored
									
										Normal file
									
												View File
												
				@@ -0,0 +1,41 @@

				#!/usr/bin/env bash

				set -eu -o pipefail

				check_tag() {

				    local expected=$1

				    local actual=$2

				    local filename=$3

				    if [[ $actual != $expected ]]; then

				        echo >&2 "Error: the current tag does not match the version in $filename: found $actual, expected $expected"

				        return 1

				    fi

				}

				read_version() {

				    grep '^version = ' | cut -d \" -f 2

				}

				if [[ -z "${GITHUB_REF:-}" ]]; then

				    echo >&2 "Error: GITHUB_REF is not set"

				    exit 1

				fi

				if [[ ! "$GITHUB_REF" =~ ^refs/tags/v[0-9]+\.[0-9]+\.[0-9]+(-[a-z0-9]+)?$ ]]; then

				    echo >&2 "Error: GITHUB_REF is not a valid tag: $GITHUB_REF"

				    exit 1

				fi

				current_tag=${GITHUB_REF#refs/tags/v}

				ret=0

				toml_tag="$(cat Cargo.toml | read_version)"

				check_tag "$current_tag" "$toml_tag" Cargo.toml || ret=1

				lock_tag=$(grep -A 1 '^name = "meilisearch-auth"' Cargo.lock | read_version)

				check_tag "$current_tag" "$lock_tag" Cargo.lock || ret=1

				if (( ret == 0 )); then

				    echo 'OK'

				fi

				exit $ret

									
										48

.github/scripts/is-latest-release.sh
									
										vendored
									
										Normal file
									
												View File
												
				@@ -0,0 +1,48 @@

				#!/bin/sh

				# Used in our CIs to publish the latest Docker image.

				# Checks if the current tag ($GITHUB_REF) corresponds to the latest release tag on GitHub

				# Returns "true" or "false" (as a string).

				GITHUB_API='https://api.github.com/repos/meilisearch/meilisearch/releases'

				PNAME='meilisearch'

				# FUNCTIONS

				# Returns the version of the latest stable version of Meilisearch by setting the $latest variable.

				get_latest() {

				    # temp_file is needed because the grep would start before the download is over

				    temp_file=$(mktemp -q /tmp/$PNAME.XXXXXXXXX)

				    latest_release="$GITHUB_API/latest"

				    if [ $? -ne 0 ]; then

				        echo "$0: Can't create temp file."

				        exit 1

				    fi

				    if [ -z "$GITHUB_PAT" ]; then

				        curl -s "$latest_release" > "$temp_file" || return 1

				    else

				        curl -H "Authorization: token $GITHUB_PAT" -s "$latest_release" > "$temp_file" || return 1

				    fi

				    latest="$(cat "$temp_file" | grep '"tag_name":' | cut -d ':' -f2 | tr -d '"' | tr -d ',' | tr -d ' ')"

				    rm -f "$temp_file"

				    return 0

				}

				# MAIN

				current_tag="$(echo $GITHUB_REF | tr -d 'refs/tags/')"

				get_latest

				if [ "$current_tag" != "$latest" ]; then

				    # The current release tag is not the latest

				    echo "false"

				else

				    # The current release tag is the latest

				    echo "true"

				fi

				exit 0

									
										20

.github/workflows/README.md
									
										vendored
									
												View File
											
				@@ -1,20 +0,0 @@

				# GitHub Actions Workflow for Meilisearch

				> **Note:**

				> - We do not use [cache](https://github.com/actions/cache) yet but we could use it to speed up CI

				## Workflow

				- On each pull request, we trigger `cargo test`.

				- On each tag, we build:

				    - the tagged Docker image and publish it to Docker Hub

				    - the binaries for MacOS, Ubuntu, and Windows

				    - the Debian package

				- On each stable release (`v*.*.*` tag):

				    - we build the `latest` Docker image and publish it to Docker Hub

				    - we publish the binary to Hombrew and Gemfury

				## Problems

				- We do not test on Windows because we are unable to make it work, there is a disk space problem.

									
										30

.github/workflows/bench-manual.yml
									
										vendored
									
										Normal file
									
												View File
												
				@@ -0,0 +1,30 @@

				name: Bench (manual)

				on:

				    workflow_dispatch:

				        inputs:

				            workload:

				                description: 'The path to the workloads to execute (workloads/...)'

				                required: true

				                default: 'workloads/movies.json'

				env:

				    WORKLOAD_NAME: ${{ github.event.inputs.workload }}

				jobs:

				    benchmarks:

				        name: Run and upload benchmarks

				        runs-on: benchmarks

				        timeout-minutes: 180 # 3h

				        steps:

				            - uses: actions/checkout@v3

				            - uses: actions-rs/toolchain@v1

				              with:

				                profile: minimal

				                toolchain: stable

				                override: true

				            - name: Run benchmarks - workload ${WORKLOAD_NAME} - branch ${{ github.ref }} - commit ${{ github.sha }}

				              run: |

				               cargo xtask bench --api-key "${{ secrets.BENCHMARK_API_KEY }}" --dashboard-url "${{ vars.BENCHMARK_DASHBOARD_URL }}" --reason "Manual [Run #${{ github.run_id }}](https://github.com/meilisearch/meilisearch/actions/runs/${{ github.run_id }})" -- ${WORKLOAD_NAME}

									
										46

.github/workflows/bench-pr.yml
									
										vendored
									
										Normal file
									
												View File
												
				@@ -0,0 +1,46 @@

				name: Bench (PR)

				on:

				    issue_comment:

				        types: [created]

				permissions:

				    issues: write

				env:

				    GH_TOKEN: ${{ secrets.MEILI_BOT_GH_PAT }}

				jobs:

				    run-benchmarks-on-comment:

				      if: startsWith(github.event.comment.body, '/bench')

				      name: Run and upload benchmarks

				      runs-on: benchmarks

				      timeout-minutes: 180 # 3h

				      steps:

				        - name: Check for Command

				          id: command

				          uses: xt0rted/slash-command-action@v2

				          with:

				              command: bench

				              reaction-type: "rocket"

				              repo-token: ${{ env.GH_TOKEN }}

				        - uses: xt0rted/pull-request-comment-branch@v2

				          id: comment-branch

				          with:

				            repo_token: ${{ env.GH_TOKEN }}

				        - uses: actions/checkout@v3

				          if: success()

				          with:

				            fetch-depth: 0 # fetch full history to be able to get main commit sha

				            ref: ${{ steps.comment-branch.outputs.head_ref }}

				        - uses: actions-rs/toolchain@v1

				          with:

				            profile: minimal

				            toolchain: stable

				            override: true

				        - name: Run benchmarks on PR ${{ github.event.issue.id }}

				          run: |

				            cargo xtask bench --api-key "${{ secrets.BENCHMARK_API_KEY }}" --dashboard-url "${{ vars.BENCHMARK_DASHBOARD_URL }}" --reason "[Comment](${{ github.event.comment.url }}) on [#${{github.event.issue.id}}](${{ github.event.issue.url }})" -- ${{ steps.command.outputs.command-arguments }}

									
										25

.github/workflows/bench-push-indexing.yml
									
										vendored
									
										Normal file
									
												View File
												
				@@ -0,0 +1,25 @@

				name: Indexing bench (push)

				on:

				    push:

				        branches:

				            - main

				jobs:

				    benchmarks:

				        name: Run and upload benchmarks

				        runs-on: benchmarks

				        timeout-minutes: 180 # 3h

				        steps:

				          - uses: actions/checkout@v3

				          - uses: actions-rs/toolchain@v1

				            with:

				              profile: minimal

				              toolchain: stable

				              override: true

				          # Run benchmarks

				          - name: Run benchmarks - Dataset ${BENCH_NAME} - Branch main - Commit ${{ github.sha }}

				            run: |

				              cargo xtask bench --api-key "${{ secrets.BENCHMARK_API_KEY }}" --dashboard-url "${{ vars.BENCHMARK_DASHBOARD_URL }}" --reason "Push on `main` [Run #${{ github.run_id }}](https://github.com/meilisearch/meilisearch/actions/runs/${{ github.run_id }})" -- workloads/*.json

									
										77

.github/workflows/benchmarks-manual.yml
									
										vendored
									
										Normal file
									
												View File
												
				@@ -0,0 +1,77 @@

				name: Benchmarks (manual)

				on:

				  workflow_dispatch:

				    inputs:

				      dataset_name:

				        description: 'The name of the dataset used to benchmark (search_songs, search_wiki, search_geo or indexing)'

				        required: false

				        default: 'search_songs'

				env:

				  BENCH_NAME: ${{ github.event.inputs.dataset_name }}

				jobs:

				  benchmarks:

				    name: Run and upload benchmarks

				    runs-on: benchmarks

				    timeout-minutes: 4320 # 72h

				    steps:

				      - uses: actions/checkout@v3

				      - uses: actions-rs/toolchain@v1

				        with:

				          profile: minimal

				          toolchain: stable

				          override: true

				      # Set variables

				      - name: Set current branch name

				        shell: bash

				        run: echo "name=$(echo ${GITHUB_REF#refs/heads/})" >> $GITHUB_OUTPUT

				        id: current_branch

				      - name: Set normalized current branch name # Replace `/` by `_` in branch name to avoid issues when pushing to S3

				        shell: bash

				        run: echo "name=$(echo ${GITHUB_REF#refs/heads/} | tr '/' '_')" >> $GITHUB_OUTPUT

				        id: normalized_current_branch

				      - name: Set shorter commit SHA

				        shell: bash

				        run: echo "short=$(echo $GITHUB_SHA | cut -c1-8)" >> $GITHUB_OUTPUT

				        id: commit_sha

				      - name: Set file basename with format "dataset_branch_commitSHA"

				        shell: bash

				        run: echo "basename=$(echo ${BENCH_NAME}_${{ steps.normalized_current_branch.outputs.name }}_${{ steps.commit_sha.outputs.short }})" >> $GITHUB_OUTPUT

				        id: file

				      # Run benchmarks

				      - name: Run benchmarks - Dataset ${BENCH_NAME} - Branch ${{ steps.current_branch.outputs.name }} - Commit ${{ steps.commit_sha.outputs.short }}

				        run: |

				          cd benchmarks

				          cargo bench --bench ${BENCH_NAME} -- --save-baseline ${{ steps.file.outputs.basename }}

				      # Generate critcmp files

				      - name: Install critcmp

				        uses: taiki-e/install-action@v2

				        with:

				          tool: critcmp

				      - name: Export cripcmp file

				        run: |

				          critcmp --export ${{ steps.file.outputs.basename }} > ${{ steps.file.outputs.basename }}.json

				      # Upload benchmarks

				      - name: Upload ${{ steps.file.outputs.basename }}.json to DO Spaces # DigitalOcean Spaces = S3

				        uses: BetaHuhn/do-spaces-action@v2

				        with:

				          access_key: ${{ secrets.DO_SPACES_ACCESS_KEY }}

				          secret_key: ${{ secrets.DO_SPACES_SECRET_KEY }}

				          space_name: ${{ secrets.DO_SPACES_SPACE_NAME }}

				          space_region: ${{ secrets.DO_SPACES_SPACE_REGION }}

				          source: ${{ steps.file.outputs.basename }}.json

				          out_dir: critcmp_results

				      # Helper

				      - name: 'README: compare with another benchmark'

				        run: |

				          echo "${{ steps.file.outputs.basename }}.json has just been pushed."

				          echo 'How to compare this benchmark with another one?'

				          echo '  - Check the available files with: ./benchmarks/scripts/list.sh'

				          echo "  - Run the following command: ./benchmaks/scripts/compare.sh <file-to-compare-with> ${{ steps.file.outputs.basename }}.json"

									
										98

.github/workflows/benchmarks-pr.yml
									
										vendored
									
										Normal file
									
												View File
												
				@@ -0,0 +1,98 @@

				name: Benchmarks (PR)

				on: issue_comment

				permissions:

				  issues: write

				env:

				  GH_TOKEN: ${{ secrets.MEILI_BOT_GH_PAT }}

				jobs:

				  run-benchmarks-on-comment:

				    if: startsWith(github.event.comment.body, '/benchmark')

				    name: Run and upload benchmarks

				    runs-on: benchmarks

				    timeout-minutes: 4320 # 72h

				    steps:

				      - uses: actions-rs/toolchain@v1

				        with:

				          profile: minimal

				          toolchain: stable

				          override: true

				      - name: Check for Command

				        id: command

				        uses: xt0rted/slash-command-action@v2

				        with:

				          command: benchmark

				          reaction-type: "eyes"

				          repo-token: ${{ env.GH_TOKEN }}

				      - uses: xt0rted/pull-request-comment-branch@v2

				        id: comment-branch

				        with:

				          repo_token: ${{ env.GH_TOKEN }}

				      - uses: actions/checkout@v3

				        if: success()

				        with:

				          fetch-depth: 0 # fetch full history to be able to get main commit sha

				          ref: ${{ steps.comment-branch.outputs.head_ref }}

				      # Set variables

				      - name: Set current branch name

				        shell: bash

				        run: echo "name=$(git rev-parse --abbrev-ref HEAD)" >> $GITHUB_OUTPUT

				        id: current_branch

				      - name: Set normalized current branch name # Replace `/` by `_` in branch name to avoid issues when pushing to S3

				        shell: bash

				        run: echo "name=$(git rev-parse --abbrev-ref HEAD | tr '/' '_')" >> $GITHUB_OUTPUT

				        id: normalized_current_branch

				      - name: Set shorter commit SHA

				        shell: bash

				        run: echo "short=$(echo $GITHUB_SHA | cut -c1-8)" >> $GITHUB_OUTPUT

				        id: commit_sha

				      - name: Set file basename with format "dataset_branch_commitSHA"

				        shell: bash

				        run: echo "basename=$(echo ${{ steps.command.outputs.command-arguments }}_${{ steps.normalized_current_branch.outputs.name }}_${{ steps.commit_sha.outputs.short }})" >> $GITHUB_OUTPUT

				        id: file

				      # Run benchmarks

				      - name: Run benchmarks - Dataset ${{ steps.command.outputs.command-arguments }} - Branch ${{ steps.current_branch.outputs.name }} - Commit ${{ steps.commit_sha.outputs.short }}

				        run: |

				          cd benchmarks

				          cargo bench --bench ${{ steps.command.outputs.command-arguments }} -- --save-baseline ${{ steps.file.outputs.basename }}

				      # Generate critcmp files

				      - name: Install critcmp

				        uses: taiki-e/install-action@v2

				        with:

				          tool: critcmp

				      - name: Export cripcmp file

				        run: |

				          critcmp --export ${{ steps.file.outputs.basename }} > ${{ steps.file.outputs.basename }}.json

				      # Upload benchmarks

				      - name: Upload ${{ steps.file.outputs.basename }}.json to DO Spaces # DigitalOcean Spaces = S3

				        uses: BetaHuhn/do-spaces-action@v2

				        with:

				          access_key: ${{ secrets.DO_SPACES_ACCESS_KEY }}

				          secret_key: ${{ secrets.DO_SPACES_SECRET_KEY }}

				          space_name: ${{ secrets.DO_SPACES_SPACE_NAME }}

				          space_region: ${{ secrets.DO_SPACES_SPACE_REGION }}

				          source: ${{ steps.file.outputs.basename }}.json

				          out_dir: critcmp_results

				      # Compute the diff of the benchmarks and send a message on the GitHub PR

				      - name: Compute and send a message in the PR

				        env:

				          GITHUB_TOKEN: ${{ secrets.MEILI_BOT_GH_PAT }}

				        run: |

				          set -x

				          export base_ref=$(git merge-base origin/main ${{ steps.comment-branch.outputs.head_ref }} | head -c8)

				          export base_filename=$(echo ${{ steps.command.outputs.command-arguments }}_main_${base_ref}.json)

				          export bench_name=$(echo ${{ steps.command.outputs.command-arguments }})

				          echo "Here are your $bench_name benchmarks diff 👊" >> body.txt

				          echo '```' >> body.txt

				          ./benchmarks/scripts/compare.sh $base_filename ${{ steps.file.outputs.basename }}.json >> body.txt

				          echo '```' >> body.txt

				          gh pr comment ${{ steps.current_branch.outputs.name }} --body-file body.txt

									
										79

.github/workflows/benchmarks-push-indexing.yml
									
										vendored
									
										Normal file
									
												View File
												
				@@ -0,0 +1,79 @@

				name: Benchmarks of indexing (push)

				on:

				  push:

				    branches:

				      - main

				env:

				  INFLUX_TOKEN: ${{ secrets.INFLUX_TOKEN }}

				  BENCH_NAME: "indexing"

				jobs:

				  benchmarks:

				    name: Run and upload benchmarks

				    runs-on: benchmarks

				    timeout-minutes: 4320 # 72h

				    steps:

				      - uses: actions/checkout@v3

				      - uses: actions-rs/toolchain@v1

				        with:

				          profile: minimal

				          toolchain: stable

				          override: true

				      # Set variables

				      - name: Set current branch name

				        shell: bash

				        run: echo "name=$(echo ${GITHUB_REF#refs/heads/})" >> $GITHUB_OUTPUT

				        id: current_branch

				      - name: Set normalized current branch name # Replace `/` by `_` in branch name to avoid issues when pushing to S3

				        shell: bash

				        run: echo "name=$(echo ${GITHUB_REF#refs/heads/} | tr '/' '_')" >> $GITHUB_OUTPUT

				        id: normalized_current_branch

				      - name: Set shorter commit SHA

				        shell: bash

				        run: echo "short=$(echo $GITHUB_SHA | cut -c1-8)" >> $GITHUB_OUTPUT

				        id: commit_sha

				      - name: Set file basename with format "dataset_branch_commitSHA"

				        shell: bash

				        run: echo "basename=$(echo ${BENCH_NAME}_${{ steps.normalized_current_branch.outputs.name }}_${{ steps.commit_sha.outputs.short }})" >> $GITHUB_OUTPUT

				        id: file

				      # Run benchmarks

				      - name: Run benchmarks - Dataset ${BENCH_NAME} - Branch ${{ steps.current_branch.outputs.name }} - Commit ${{ steps.commit_sha.outputs.short }}

				        run: |

				          cd benchmarks

				          cargo bench --bench ${BENCH_NAME} -- --save-baseline ${{ steps.file.outputs.basename }}

				      # Generate critcmp files

				      - name: Install critcmp

				        uses: taiki-e/install-action@v2

				        with:

				          tool: critcmp

				      - name: Export cripcmp file

				        run: |

				          critcmp --export ${{ steps.file.outputs.basename }} > ${{ steps.file.outputs.basename }}.json

				      # Upload benchmarks

				      - name: Upload ${{ steps.file.outputs.basename }}.json to DO Spaces # DigitalOcean Spaces = S3

				        uses: BetaHuhn/do-spaces-action@v2

				        with:

				          access_key: ${{ secrets.DO_SPACES_ACCESS_KEY }}

				          secret_key: ${{ secrets.DO_SPACES_SECRET_KEY }}

				          space_name: ${{ secrets.DO_SPACES_SPACE_NAME }}

				          space_region: ${{ secrets.DO_SPACES_SPACE_REGION }}

				          source: ${{ steps.file.outputs.basename }}.json

				          out_dir: critcmp_results

				      # Upload benchmarks to influxdb

				      - name: Upload ${{ steps.file.outputs.basename }}.json to influxDB

				        run: telegraf --config https://eu-central-1-1.aws.cloud2.influxdata.com/api/v2/telegrafs/08b52e34a370b000 --once --debug

				      # Helper

				      - name: 'README: compare with another benchmark'

				        run: |

				          echo "${{ steps.file.outputs.basename }}.json has just been pushed."

				          echo 'How to compare this benchmark with another one?'

				          echo '  - Check the available files with: ./benchmarks/scripts/list.sh'

				          echo "  - Run the following command: ./benchmaks/scipts/compare.sh <file-to-compare-with> ${{ steps.file.outputs.basename }}.json"

									
										78

.github/workflows/benchmarks-push-search-geo.yml
									
										vendored
									
										Normal file
									
												View File
												
				@@ -0,0 +1,78 @@

				name: Benchmarks of search for geo (push)

				on:

				  push:

				    branches:

				      - main

				env:

				  BENCH_NAME: "search_geo"

				  INFLUX_TOKEN: ${{ secrets.INFLUX_TOKEN }}

				jobs:

				  benchmarks:

				    name: Run and upload benchmarks

				    runs-on: benchmarks

				    steps:

				      - uses: actions/checkout@v3

				      - uses: actions-rs/toolchain@v1

				        with:

				          profile: minimal

				          toolchain: stable

				          override: true

				      # Set variables

				      - name: Set current branch name

				        shell: bash

				        run: echo "name=$(echo ${GITHUB_REF#refs/heads/})" >> $GITHUB_OUTPUT

				        id: current_branch

				      - name: Set normalized current branch name # Replace `/` by `_` in branch name to avoid issues when pushing to S3

				        shell: bash

				        run: echo "name=$(echo ${GITHUB_REF#refs/heads/} | tr '/' '_')" >> $GITHUB_OUTPUT

				        id: normalized_current_branch

				      - name: Set shorter commit SHA

				        shell: bash

				        run: echo "short=$(echo $GITHUB_SHA | cut -c1-8)" >> $GITHUB_OUTPUT

				        id: commit_sha

				      - name: Set file basename with format "dataset_branch_commitSHA"

				        shell: bash

				        run: echo "basename=$(echo ${BENCH_NAME}_${{ steps.normalized_current_branch.outputs.name }}_${{ steps.commit_sha.outputs.short }})" >> $GITHUB_OUTPUT

				        id: file

				      # Run benchmarks

				      - name: Run benchmarks - Dataset ${BENCH_NAME} - Branch ${{ steps.current_branch.outputs.name }} - Commit ${{ steps.commit_sha.outputs.short }}

				        run: |

				          cd benchmarks

				          cargo bench --bench ${BENCH_NAME} -- --save-baseline ${{ steps.file.outputs.basename }}

				      # Generate critcmp files

				      - name: Install critcmp

				        uses: taiki-e/install-action@v2

				        with:

				          tool: critcmp

				      - name: Export cripcmp file

				        run: |

				          critcmp --export ${{ steps.file.outputs.basename }} > ${{ steps.file.outputs.basename }}.json

				      # Upload benchmarks

				      - name: Upload ${{ steps.file.outputs.basename }}.json to DO Spaces # DigitalOcean Spaces = S3

				        uses: BetaHuhn/do-spaces-action@v2

				        with:

				          access_key: ${{ secrets.DO_SPACES_ACCESS_KEY }}

				          secret_key: ${{ secrets.DO_SPACES_SECRET_KEY }}

				          space_name: ${{ secrets.DO_SPACES_SPACE_NAME }}

				          space_region: ${{ secrets.DO_SPACES_SPACE_REGION }}

				          source: ${{ steps.file.outputs.basename }}.json

				          out_dir: critcmp_results

				      # Upload benchmarks to influxdb

				      - name: Upload ${{ steps.file.outputs.basename }}.json to influxDB

				        run: telegraf --config https://eu-central-1-1.aws.cloud2.influxdata.com/api/v2/telegrafs/08b52e34a370b000 --once --debug

				      # Helper

				      - name: 'README: compare with another benchmark'

				        run: |

				          echo "${{ steps.file.outputs.basename }}.json has just been pushed."

				          echo 'How to compare this benchmark with another one?'

				          echo '  - Check the available files with: ./benchmarks/scripts/list.sh'

				          echo "  - Run the following command: ./benchmaks/scipts/compare.sh <file-to-compare-with> ${{ steps.file.outputs.basename }}.json"

									
										78

.github/workflows/benchmarks-push-search-songs.yml
									
										vendored
									
										Normal file
									
												View File
												
				@@ -0,0 +1,78 @@

				name: Benchmarks of search for songs (push)

				on:

				  push:

				    branches:

				      - main

				env:

				  BENCH_NAME: "search_songs"

				  INFLUX_TOKEN: ${{ secrets.INFLUX_TOKEN }}

				jobs:

				  benchmarks:

				    name: Run and upload benchmarks

				    runs-on: benchmarks

				    steps:

				      - uses: actions/checkout@v3

				      - uses: actions-rs/toolchain@v1

				        with:

				          profile: minimal

				          toolchain: stable

				          override: true

				      # Set variables

				      - name: Set current branch name

				        shell: bash

				        run: echo "name=$(echo ${GITHUB_REF#refs/heads/})" >> $GITHUB_OUTPUT

				        id: current_branch

				      - name: Set normalized current branch name # Replace `/` by `_` in branch name to avoid issues when pushing to S3

				        shell: bash

				        run: echo "name=$(echo ${GITHUB_REF#refs/heads/} | tr '/' '_')" >> $GITHUB_OUTPUT

				        id: normalized_current_branch

				      - name: Set shorter commit SHA

				        shell: bash

				        run: echo "short=$(echo $GITHUB_SHA | cut -c1-8)" >> $GITHUB_OUTPUT

				        id: commit_sha

				      - name: Set file basename with format "dataset_branch_commitSHA"

				        shell: bash

				        run: echo "basename=$(echo ${BENCH_NAME}_${{ steps.normalized_current_branch.outputs.name }}_${{ steps.commit_sha.outputs.short }})" >> $GITHUB_OUTPUT

				        id: file

				      # Run benchmarks

				      - name: Run benchmarks - Dataset ${BENCH_NAME} - Branch ${{ steps.current_branch.outputs.name }} - Commit ${{ steps.commit_sha.outputs.short }}

				        run: |

				          cd benchmarks

				          cargo bench --bench ${BENCH_NAME} -- --save-baseline ${{ steps.file.outputs.basename }}

				      # Generate critcmp files

				      - name: Install critcmp

				        uses: taiki-e/install-action@v2

				        with:

				          tool: critcmp

				      - name: Export cripcmp file

				        run: |

				          critcmp --export ${{ steps.file.outputs.basename }} > ${{ steps.file.outputs.basename }}.json

				      # Upload benchmarks

				      - name: Upload ${{ steps.file.outputs.basename }}.json to DO Spaces # DigitalOcean Spaces = S3

				        uses: BetaHuhn/do-spaces-action@v2

				        with:

				          access_key: ${{ secrets.DO_SPACES_ACCESS_KEY }}

				          secret_key: ${{ secrets.DO_SPACES_SECRET_KEY }}

				          space_name: ${{ secrets.DO_SPACES_SPACE_NAME }}

				          space_region: ${{ secrets.DO_SPACES_SPACE_REGION }}

				          source: ${{ steps.file.outputs.basename }}.json

				          out_dir: critcmp_results

				      # Upload benchmarks to influxdb

				      - name: Upload ${{ steps.file.outputs.basename }}.json to influxDB

				        run: telegraf --config https://eu-central-1-1.aws.cloud2.influxdata.com/api/v2/telegrafs/08b52e34a370b000 --once --debug

				      # Helper

				      - name: 'README: compare with another benchmark'

				        run: |

				          echo "${{ steps.file.outputs.basename }}.json has just been pushed."

				          echo 'How to compare this benchmark with another one?'

				          echo '  - Check the available files with: ./benchmarks/scripts/list.sh'

				          echo "  - Run the following command: ./benchmaks/scipts/compare.sh <file-to-compare-with> ${{ steps.file.outputs.basename }}.json"

									
										78

.github/workflows/benchmarks-push-search-wiki.yml
									
										vendored
									
										Normal file
									
												View File
												
				@@ -0,0 +1,78 @@

				name: Benchmarks of search for Wikipedia articles (push)

				on:

				  push:

				    branches:

				      - main

				env:

				  BENCH_NAME: "search_wiki"

				  INFLUX_TOKEN: ${{ secrets.INFLUX_TOKEN }}

				jobs:

				  benchmarks:

				    name: Run and upload benchmarks

				    runs-on: benchmarks

				    steps:

				      - uses: actions/checkout@v3

				      - uses: actions-rs/toolchain@v1

				        with:

				          profile: minimal

				          toolchain: stable

				          override: true

				      # Set variables

				      - name: Set current branch name

				        shell: bash

				        run: echo "name=$(echo ${GITHUB_REF#refs/heads/})" >> $GITHUB_OUTPUT

				        id: current_branch

				      - name: Set normalized current branch name # Replace `/` by `_` in branch name to avoid issues when pushing to S3

				        shell: bash

				        run: echo "name=$(echo ${GITHUB_REF#refs/heads/} | tr '/' '_')" >> $GITHUB_OUTPUT

				        id: normalized_current_branch

				      - name: Set shorter commit SHA

				        shell: bash

				        run: echo "short=$(echo $GITHUB_SHA | cut -c1-8)" >> $GITHUB_OUTPUT

				        id: commit_sha

				      - name: Set file basename with format "dataset_branch_commitSHA"

				        shell: bash

				        run: echo "basename=$(echo ${BENCH_NAME}_${{ steps.normalized_current_branch.outputs.name }}_${{ steps.commit_sha.outputs.short }})" >> $GITHUB_OUTPUT

				        id: file

				      # Run benchmarks

				      - name: Run benchmarks - Dataset ${BENCH_NAME} - Branch ${{ steps.current_branch.outputs.name }} - Commit ${{ steps.commit_sha.outputs.short }}

				        run: |

				          cd benchmarks

				          cargo bench --bench ${BENCH_NAME} -- --save-baseline ${{ steps.file.outputs.basename }}

				      # Generate critcmp files

				      - name: Install critcmp

				        uses: taiki-e/install-action@v2

				        with:

				          tool: critcmp

				      - name: Export cripcmp file

				        run: |

				          critcmp --export ${{ steps.file.outputs.basename }} > ${{ steps.file.outputs.basename }}.json

				      # Upload benchmarks

				      - name: Upload ${{ steps.file.outputs.basename }}.json to DO Spaces # DigitalOcean Spaces = S3

				        uses: BetaHuhn/do-spaces-action@v2

				        with:

				          access_key: ${{ secrets.DO_SPACES_ACCESS_KEY }}

				          secret_key: ${{ secrets.DO_SPACES_SECRET_KEY }}

				          space_name: ${{ secrets.DO_SPACES_SPACE_NAME }}

				          space_region: ${{ secrets.DO_SPACES_SPACE_REGION }}

				          source: ${{ steps.file.outputs.basename }}.json

				          out_dir: critcmp_results

				      # Upload benchmarks to influxdb

				      - name: Upload ${{ steps.file.outputs.basename }}.json to influxDB

				        run: telegraf --config https://eu-central-1-1.aws.cloud2.influxdata.com/api/v2/telegrafs/08b52e34a370b000 --once --debug

				      # Helper

				      - name: 'README: compare with another benchmark'

				        run: |

				          echo "${{ steps.file.outputs.basename }}.json has just been pushed."

				          echo 'How to compare this benchmark with another one?'

				          echo '  - Check the available files with: ./benchmarks/scripts/list.sh'

				          echo "  - Run the following command: ./benchmaks/scipts/compare.sh <file-to-compare-with> ${{ steps.file.outputs.basename }}.json"

									
										33

.github/workflows/coverage.yml
									
										vendored
									
												View File
											
				@@ -1,33 +0,0 @@

				---

				on:

				  workflow_dispatch:

				name: Execute code coverage

				jobs:

				  nightly-coverage:

				    runs-on: ubuntu-18.04

				    steps:

				      - uses: actions/checkout@v2

				      - uses: actions-rs/toolchain@v1

				        with:

				          toolchain: nightly

				          override: true

				      - uses: actions-rs/cargo@v1

				        with:

				          command: clean

				      - uses: actions-rs/cargo@v1

				        with:

				          command: test

				          args: --all-features --no-fail-fast

				        env:

				          CARGO_INCREMENTAL: "0"

				          RUSTFLAGS: "-Zprofile -Ccodegen-units=1 -Cinline-threshold=0 -Clink-dead-code -Coverflow-checks=off -Cpanic=unwind -Zpanic_abort_tests"

				      - uses: actions-rs/grcov@v0.1

				      - name: Upload coverage to Codecov

				        uses: codecov/codecov-action@v1

				        with:

				          token: ${{ secrets.CODECOV_TOKEN }}

				          file: ${{ steps.coverage.outputs.report }}

				          yml: ./codecov.yml

				          fail_ci_if_error: true

									
										24

.github/workflows/dependency-issue.yml
									
										vendored
									
										Normal file
									
												View File
												
				@@ -0,0 +1,24 @@

				name: Create issue to upgrade dependencies

				on:

				  schedule:

				    # Run the first of the month, every 6 month

				    - cron: '0 0 1 */6 *'

				  workflow_dispatch:

				jobs:

				  create-issue:

				    runs-on: ubuntu-latest

				    env:

				      ISSUE_TEMPLATE: issue-template.md

				      GH_TOKEN: ${{ secrets.MEILI_BOT_GH_PAT }}

				    steps:

				    - uses: actions/checkout@v3

				    - name: Download the issue template

				      run: curl -s https://raw.githubusercontent.com/meilisearch/engine-team/main/issue-templates/dependency-issue.md > $ISSUE_TEMPLATE

				    - name: Create issue

				      run: |

				        gh issue create \

				          --title 'Upgrade dependencies' \

				          --label 'dependencies,maintenance' \

				          --body-file $ISSUE_TEMPLATE

									
										32

.github/workflows/flaky-tests.yml
									
										vendored
									
										Normal file
									
												View File
												
				@@ -0,0 +1,32 @@

				name: Look for flaky tests

				on:

				  workflow_dispatch:

				  schedule:

				    - cron: "0 12 * * FRI" # Every Friday at 12:00PM

				jobs:

				  flaky:

				    runs-on: ubuntu-latest

				    container:

				      # Use ubuntu-18.04 to compile with glibc 2.27, which are the production expectations

				      image: ubuntu:18.04

				    steps:

				    - uses: actions/checkout@v3

				    - name: Install needed dependencies

				      run: |

				        apt-get update && apt-get install -y curl

				        apt-get install build-essential -y

				    - uses: actions-rs/toolchain@v1

				      with:

				        toolchain: stable

				        override: true

				    - name: Install cargo-flaky

				      run: cargo install cargo-flaky

				    - name: Run cargo flaky in the dumps

				      run: cd dump; cargo flaky -i 100 --release

				    - name: Run cargo flaky in the index-scheduler

				      run: cd index-scheduler; cargo flaky -i 100 --release

				    - name: Run cargo flaky in the auth

				      run: cd meilisearch-auth; cargo flaky -i 100 --release

				    - name: Run cargo flaky in meilisearch

				      run: cd meilisearch; cargo flaky -i 100 --release

									
										15

.github/workflows/flaky.yml
									
										vendored
									
												View File
											
				@@ -1,15 +0,0 @@

				name: Look for flaky tests

				on:

				  schedule:

				    - cron: "0 12 * * FRI" # every friday at 12:00PM

				jobs:

				  flaky:

				    runs-on: ubuntu-18.04

				    steps:

				    - uses: actions/checkout@v2

				    - name: Install cargo-flaky

				      run: cargo install cargo-flaky

				    - name: Run cargo flaky 100 times

				      run: cargo flaky -i 100 --release

									
										24

.github/workflows/fuzzer-indexing.yml
									
										vendored
									
										Normal file
									
												View File
												
				@@ -0,0 +1,24 @@

				name: Run the indexing fuzzer

				on:

				  push:

				    branches:

				      - main

				jobs:

				  fuzz:

				    name: Setup the action

				    runs-on: ubuntu-latest

				    timeout-minutes: 4320 # 72h

				    steps:

				      - uses: actions/checkout@v3

				      - uses: actions-rs/toolchain@v1

				        with:

				          profile: minimal

				          toolchain: stable

				          override: true

				      # Run benchmarks

				      - name: Run the fuzzer

				        run: |

				          cargo run --release --bin fuzz-indexing

									
										29

.github/workflows/latest-git-tag.yml
									
										vendored
									
										Normal file
									
												View File
												
				@@ -0,0 +1,29 @@

				# Create or update a latest git tag when releasing a stable version of Meilisearch

				name: Update latest git tag

				on:

				  workflow_dispatch:

				  release:

				    types: [released]

				jobs:

				  check-version:

				    name: Check the version validity

				    runs-on: ubuntu-latest

				    steps:

				      - uses: actions/checkout@v3

				      - name: Check release validity

				        if: github.event_name == 'release'

				        run: bash .github/scripts/check-release.sh

				  update-latest-tag:

				    runs-on: ubuntu-latest

				    needs: check-version

				    steps:

				      - uses: actions/checkout@v3

				      - uses: rickstaa/action-create-tag@v1

				        with:

				          tag: "latest"

				          message: "Latest stable release of Meilisearch"

				          # Move the tag if `latest` already exists

				          force_push_tag: true

				          github_token: ${{ secrets.MEILI_BOT_GH_PAT }}

									
										154

.github/workflows/milestone-workflow.yml
									
										vendored
									
										Normal file
									
												View File
												
				@@ -0,0 +1,154 @@

				name: Milestone's workflow

				# /!\ No git flow are handled here

				# For each Milestone created (not opened!), and if the release is NOT a patch release (only the patch changed)

				# - the roadmap issue is created, see https://github.com/meilisearch/engine-team/blob/main/issue-templates/roadmap-issue.md

				# - the changelog issue is created, see https://github.com/meilisearch/engine-team/blob/main/issue-templates/changelog-issue.md

				# For each Milestone closed

				# - the `release_version` label is created

				# - this label is applied to all issues/PRs in the Milestone

				on:

				  milestone:

				    types: [created, closed]

				env:

				  MILESTONE_VERSION: ${{ github.event.milestone.title }}

				  MILESTONE_URL: ${{ github.event.milestone.html_url }}

				  MILESTONE_DUE_ON: ${{ github.event.milestone.due_on }}

				  GH_TOKEN: ${{ secrets.MEILI_BOT_GH_PAT }}

				jobs:

				# -----------------

				# MILESTONE CREATED

				# -----------------

				  get-release-version:

				    if: github.event.action == 'created'

				    runs-on: ubuntu-latest

				    outputs:

				      is-patch: ${{ steps.check-patch.outputs.is-patch }}

				    steps:

				      - uses: actions/checkout@v3

				      - name: Check if this release is a patch release only

				        id: check-patch

				        run: |

				          echo version: $MILESTONE_VERSION

				          if [[ $MILESTONE_VERSION =~ ^v[0-9]+\.[0-9]+\.0$ ]]; then

				            echo 'This is NOT a patch release'

				            echo "is-patch=false" >> $GITHUB_OUTPUT

				          elif [[ $MILESTONE_VERSION =~ ^v[0-9]+\.[0-9]+\.[0-9]+$ ]]; then

				            echo 'This is a patch release'

				            echo "is-patch=true" >> $GITHUB_OUTPUT

				          else

				            echo "Not a valid format of release, check the Milestone's title."

				            echo 'Should be vX.Y.Z'

				            exit 1

				          fi

				  create-roadmap-issue:

				    needs: get-release-version

				    # Create the roadmap issue if the release is not only a patch release

				    if: github.event.action == 'created' && needs.get-release-version.outputs.is-patch == 'false'

				    runs-on: ubuntu-latest

				    env:

				      ISSUE_TEMPLATE: issue-template.md

				    steps:

				      - uses: actions/checkout@v3

				      - name: Download the issue template

				        run: curl -s https://raw.githubusercontent.com/meilisearch/engine-team/main/issue-templates/roadmap-issue.md > $ISSUE_TEMPLATE

				      - name: Replace all empty occurrences in the templates

				        run: |

				          # Replace all <<version>> occurrences

				          sed -i "s/<<version>>/$MILESTONE_VERSION/g" $ISSUE_TEMPLATE

				          # Replace all <<milestone_id>> occurrences

				          milestone_id=$(echo $MILESTONE_URL | cut -d '/' -f 7)

				          sed -i "s/<<milestone_id>>/$milestone_id/g" $ISSUE_TEMPLATE

				          # Replace release date if exists

				          if [[ ! -z $MILESTONE_DUE_ON ]]; then

				            date=$(echo $MILESTONE_DUE_ON | cut -d 'T' -f 1)

				            sed -i "s/Release date\: 20XX-XX-XX/Release date\: $date/g" $ISSUE_TEMPLATE

				          fi

				      - name: Create the issue

				        run: |

				          gh issue create \

				            --title "$MILESTONE_VERSION ROADMAP" \

				            --label 'epic,impacts docs,impacts integrations,impacts cloud' \

				            --body-file $ISSUE_TEMPLATE \

				            --milestone $MILESTONE_VERSION

				  create-changelog-issue:

				    needs: get-release-version

				    # Create the changelog issue if the release is not only a patch release

				    if: github.event.action == 'created' && needs.get-release-version.outputs.is-patch == 'false'

				    runs-on: ubuntu-latest

				    env:

				      ISSUE_TEMPLATE: issue-template.md

				    steps:

				      - uses: actions/checkout@v3

				      - name: Download the issue template

				        run: curl -s https://raw.githubusercontent.com/meilisearch/engine-team/main/issue-templates/changelog-issue.md > $ISSUE_TEMPLATE

				      - name: Replace all empty occurrences in the templates

				        run: |

				          # Replace all <<version>> occurrences

				          sed -i "s/<<version>>/$MILESTONE_VERSION/g" $ISSUE_TEMPLATE

				          # Replace all <<milestone_id>> occurrences

				          milestone_id=$(echo $MILESTONE_URL | cut -d '/' -f 7)

				          sed -i "s/<<milestone_id>>/$milestone_id/g" $ISSUE_TEMPLATE

				      - name: Create the issue

				        run: |

				          gh issue create \

				            --title "Create release changelogs for $MILESTONE_VERSION" \

				            --label 'impacts docs,documentation' \

				            --body-file $ISSUE_TEMPLATE \

				            --milestone $MILESTONE_VERSION \

				            --assignee curquiza

				# ----------------

				# MILESTONE CLOSED

				# ----------------

				  create-release-label:

				    if: github.event.action == 'closed'

				    runs-on: ubuntu-latest

				    steps:

				      - uses: actions/checkout@v3

				      - name: Create the ${{ env.MILESTONE_VERSION }} label

				        run: |

				          label_description="PRs/issues solved in $MILESTONE_VERSION"

				          if [[ ! -z $MILESTONE_DUE_ON ]]; then

				            date=$(echo $MILESTONE_DUE_ON | cut -d 'T' -f 1)

				            label_description="$label_description released on $date"

				          fi

				          gh api repos/meilisearch/meilisearch/labels \

				            --method POST \

				            -H "Accept: application/vnd.github+json" \

				            -f name="$MILESTONE_VERSION" \

				            -f description="$label_description" \

				            -f color='ff5ba3'

				  labelize-all-milestone-content:

				    if: github.event.action == 'closed'

				    needs: create-release-label

				    runs-on: ubuntu-latest

				    steps:

				      - uses: actions/checkout@v3

				      - name: Add label ${{ env.MILESTONE_VERSION }} to all PRs in the Milestone

				        run: |

				          prs=$(gh pr list --search milestone:"$MILESTONE_VERSION" --limit 1000 --state all --json number --template '{{range .}}{{tablerow (printf "%v" .number)}}{{end}}')

				          for pr in $prs; do

				              gh pr edit $pr --add-label $MILESTONE_VERSION

				          done

				      - name: Add label ${{ env.MILESTONE_VERSION }} to all issues in the Milestone

				        run: |

				          issues=$(gh issue list --search milestone:"$MILESTONE_VERSION" --limit 1000 --state all --json number --template '{{range .}}{{tablerow (printf "%v" .number)}}{{end}}')

				          for issue in $issues; do

				              gh issue edit $issue --add-label $MILESTONE_VERSION

				          done

									
										58

.github/workflows/publish-apt-brew-pkg.yml
									
										vendored
									
										Normal file
									
												View File
												
				@@ -0,0 +1,58 @@

				name: Publish to APT & Homebrew

				on:

				  release:

				    types: [released]

				jobs:

				  check-version:

				    name: Check the version validity

				    runs-on: ubuntu-latest

				    steps:

				      - uses: actions/checkout@v3

				      - name: Check release validity

				        run: bash .github/scripts/check-release.sh

				  debian:

				    name: Publish debian packagge

				    runs-on: ubuntu-latest

				    needs: check-version

				    container:

				      # Use ubuntu-18.04 to compile with glibc 2.27

				      image: ubuntu:18.04

				    steps:

				    - name: Install needed dependencies

				      run: |

				        apt-get update && apt-get install -y curl

				        apt-get install build-essential -y

				    - uses: actions-rs/toolchain@v1

				      with:

				        toolchain: stable

				        override: true

				    - name: Install cargo-deb

				      run: cargo install cargo-deb

				    - uses: actions/checkout@v3

				    - name: Build deb package

				      run: cargo deb -p meilisearch -o target/debian/meilisearch.deb

				    - name: Upload debian pkg to release

				      uses: svenstaro/upload-release-action@2.7.0

				      with:

				        repo_token: ${{ secrets.MEILI_BOT_GH_PAT }}

				        file: target/debian/meilisearch.deb

				        asset_name: meilisearch.deb

				        tag: ${{ github.ref }}

				    - name: Upload debian pkg to apt repository

				      run: curl -F package=@target/debian/meilisearch.deb https://${{ secrets.GEMFURY_PUSH_TOKEN }}@push.fury.io/meilisearch/

				  homebrew:

				    name: Bump Homebrew formula

				    runs-on: ubuntu-latest

				    needs: check-version

				    steps:

				      - name: Create PR to Homebrew

				        uses: mislav/bump-homebrew-formula-action@v3

				        with:

				          formula-name: meilisearch

				          formula-path: Formula/m/meilisearch.rb

				        env:

				          COMMITTER_TOKEN: ${{ secrets.HOMEBREW_COMMITTER_TOKEN }}

									
										173

.github/workflows/publish-binaries.yml
									
										vendored
									
												View File
												
				@@ -1,62 +1,111 @@

				name: Publish binaries to GitHub release

				on:

				  workflow_dispatch:

				  schedule:

				    - cron: '0 2 * * *' # Every day at 2:00am

				  release:

				    types: [published]

				name: Publish binaries to release

				jobs:

				  publish:

				  check-version:

				    name: Check the version validity

				    runs-on: ubuntu-latest

				    # No need to check the version for dry run (cron)

				    steps:

				      - uses: actions/checkout@v3

				      # Check if the tag has the v<nmumber>.<number>.<number> format.

				      # If yes, it means we are publishing an official release.

				      # If no, we are releasing a RC, so no need to check the version.

				      - name: Check tag format

				        if: github.event_name == 'release'

				        id: check-tag-format

				        run: |

				          escaped_tag=$(printf "%q" ${{ github.ref_name }})

				          if [[ $escaped_tag =~ ^v[0-9]+\.[0-9]+\.[0-9]+$ ]]; then

				            echo "stable=true" >> $GITHUB_OUTPUT

				          else

				            echo "stable=false" >> $GITHUB_OUTPUT

				          fi

				      - name: Check release validity

				        if: github.event_name == 'release' && steps.check-tag-format.outputs.stable == 'true'

				        run: bash .github/scripts/check-release.sh

				  publish-linux:

				    name: Publish binary for Linux

				    runs-on: ubuntu-latest

				    needs: check-version

				    container:

				      # Use ubuntu-18.04 to compile with glibc 2.27

				      image: ubuntu:18.04

				    steps:

				    - uses: actions/checkout@v3

				    - name: Install needed dependencies

				      run: |

				        apt-get update && apt-get install -y curl

				        apt-get install build-essential -y

				    - uses: actions-rs/toolchain@v1

				      with:

				        toolchain: stable

				        override: true

				    - name: Build

				      run: cargo build --release --locked

				    # No need to upload binaries for dry run (cron)

				    - name: Upload binaries to release

				      if: github.event_name == 'release'

				      uses: svenstaro/upload-release-action@2.7.0

				      with:

				        repo_token: ${{ secrets.MEILI_BOT_GH_PAT }}

				        file: target/release/meilisearch

				        asset_name: meilisearch-linux-amd64

				        tag: ${{ github.ref }}

				  publish-macos-windows:

				    name: Publish binary for ${{ matrix.os }}

				    runs-on: ${{ matrix.os }}

				    needs: check-version

				    strategy:

				      fail-fast: false

				      matrix:

				        os: [ubuntu-18.04, macos-latest, windows-latest]

				        os: [macos-12, windows-2022]

				        include:

				          - os: ubuntu-18.04

				            artifact_name: meilisearch

				            asset_name: meilisearch-linux-amd64

				          - os: macos-latest

				          - os: macos-12

				            artifact_name: meilisearch

				            asset_name: meilisearch-macos-amd64

				          - os: windows-latest

				          - os: windows-2022

				            artifact_name: meilisearch.exe

				            asset_name: meilisearch-windows-amd64.exe

				    steps:

				    - uses: hecrj/setup-rust-action@master

				    - uses: actions/checkout@v3

				    - uses: actions-rs/toolchain@v1

				      with:

				        rust-version: stable

				    - uses: actions/checkout@v2

				        toolchain: stable

				        override: true

				    - name: Build

				      run: cargo build --release --locked

				    # No need to upload binaries for dry run (cron)

				    - name: Upload binaries to release

				      uses: svenstaro/upload-release-action@v1-release

				      if: github.event_name == 'release'

				      uses: svenstaro/upload-release-action@2.7.0

				      with:

				        repo_token: ${{ secrets.PUBLISH_TOKEN }}

				        repo_token: ${{ secrets.MEILI_BOT_GH_PAT }}

				        file: target/release/${{ matrix.artifact_name }}

				        asset_name: ${{ matrix.asset_name }}

				        tag: ${{ github.ref }}

				  publish-aarch64:

				    name: Publish binary for aarch64

				    runs-on: ${{ matrix.os }}

				    continue-on-error: false

				  publish-macos-apple-silicon:

				    name: Publish binary for macOS silicon

				    runs-on: macos-12

				    needs: check-version

				    strategy:

				      fail-fast: false

				      matrix:

				        include:

				          - build: aarch64

				            os: ubuntu-18.04

				            target: aarch64-unknown-linux-gnu

				            linker: gcc-aarch64-linux-gnu

				            use-cross: true

				            asset_name: meilisearch-linux-aarch64

				          - target: aarch64-apple-darwin

				            asset_name: meilisearch-macos-apple-silicon

				    steps:

				      - name: Checkout repository

				        uses: actions/checkout@v2

				        uses: actions/checkout@v3

				      - name: Installing Rust toolchain

				        uses: actions-rs/toolchain@v1

				        with:

				@@ -64,18 +113,54 @@ jobs:

				          profile: minimal

				          target: ${{ matrix.target }}

				          override: true

				      - name: Cargo build

				        uses: actions-rs/cargo@v1

				        with:

				          command: build

				          args: --release --target ${{ matrix.target }}

				      - name: Upload the binary to release

				        # No need to upload binaries for dry run (cron)

				        if: github.event_name == 'release'

				        uses: svenstaro/upload-release-action@2.7.0

				        with:

				          repo_token: ${{ secrets.MEILI_BOT_GH_PAT }}

				          file: target/${{ matrix.target }}/release/meilisearch

				          asset_name: ${{ matrix.asset_name }}

				          tag: ${{ github.ref }}

				      - name: APT update

				  publish-aarch64:

				    name: Publish binary for aarch64

				    runs-on: ubuntu-latest

				    needs: check-version

				    container:

				      # Use ubuntu-18.04 to compile with glibc 2.27

				      image: ubuntu:18.04

				    strategy:

				      matrix:

				        include:

				          - target: aarch64-unknown-linux-gnu

				            asset_name: meilisearch-linux-aarch64

				    steps:

				      - name: Checkout repository

				        uses: actions/checkout@v3

				      - name: Install needed dependencies

				        run: |

				          sudo apt update

				      - name: Install target specific tools

				        if: matrix.use-cross

				          apt-get update -y && apt upgrade -y

				          apt-get install -y curl build-essential gcc-aarch64-linux-gnu

				      - name: Set up Docker for cross compilation

				        run: |

				          sudo apt-get install -y ${{ matrix.linker }}

				          apt-get install -y curl apt-transport-https ca-certificates software-properties-common

				          curl -fsSL https://download.docker.com/linux/ubuntu/gpg | apt-key add -

				          add-apt-repository "deb [arch=$(dpkg --print-architecture)] https://download.docker.com/linux/ubuntu $(lsb_release -cs) stable"

				          apt-get update -y && apt-get install -y docker-ce

				      - name: Installing Rust toolchain

				        uses: actions-rs/toolchain@v1

				        with:

				          toolchain: stable

				          profile: minimal

				          target: ${{ matrix.target }}

				          override: true

				      - name: Configure target aarch64 GNU

				        if: matrix.target == 'aarch64-unknown-linux-gnu'

				        ## Environment variable is not passed using env:

				        ## LD gold won't work with MUSL

				        # env:

				@@ -85,22 +170,22 @@ jobs:

				          echo '[target.aarch64-unknown-linux-gnu]' >> ~/.cargo/config

				          echo 'linker = "aarch64-linux-gnu-gcc"' >> ~/.cargo/config

				          echo 'JEMALLOC_SYS_WITH_LG_PAGE=16' >> $GITHUB_ENV

				          echo RUSTFLAGS="-Clink-arg=-fuse-ld=gold" >> $GITHUB_ENV

				      - name: Cargo build

				        uses: actions-rs/cargo@v1

				        with:

				          command: build

				          use-cross: ${{ matrix.use-cross }}

				          use-cross: true

				          args: --release --target ${{ matrix.target }}

				        env:

				          CROSS_DOCKER_IN_DOCKER: true

				      - name: List target output files

				        run: ls -lR ./target

				      - name: Upload the binary to release

				        uses: svenstaro/upload-release-action@v1-release

				        # No need to upload binaries for dry run (cron)

				        if: github.event_name == 'release'

				        uses: svenstaro/upload-release-action@2.7.0

				        with:

				          repo_token: ${{ secrets.PUBLISH_TOKEN }}

				          repo_token: ${{ secrets.MEILI_BOT_GH_PAT }}

				          file: target/${{ matrix.target }}/release/meilisearch

				          asset_name: ${{ matrix.asset_name }}

				          tag: ${{ github.ref }}

									
										39

.github/workflows/publish-deb-brew-pkg.yml
									
										vendored
									
												View File
											
				@@ -1,39 +0,0 @@

				name: Publish deb pkg to GitHub release & APT repository & Homebrew

				on:

				  release:

				    types: [released]

				jobs:

				  debian:

				    name: Publish debian packagge

				    runs-on: ubuntu-18.04

				    steps:

				    - uses: hecrj/setup-rust-action@master

				      with:

				        rust-version: stable

				    - name: Install cargo-deb

				      run: cargo install cargo-deb

				    - uses: actions/checkout@v2

				    - name: Build deb package

				      run: cargo deb -p meilisearch-http -o target/debian/meilisearch.deb

				    - name: Upload debian pkg to release

				      uses: svenstaro/upload-release-action@v1-release

				      with:

				        repo_token: ${{ secrets.GITHUB_TOKEN }}

				        file: target/debian/meilisearch.deb

				        asset_name: meilisearch.deb

				        tag: ${{ github.ref }}

				    - name: Upload debian pkg to apt repository

				      run: curl -F package=@target/debian/meilisearch.deb https://${{ secrets.GEMFURY_PUSH_TOKEN }}@push.fury.io/meilisearch/

				  homebrew:

				    name: Bump Homebrew formula

				    runs-on: ubuntu-18.04

				    steps:

				      - name: Create PR to Homebrew

				        uses: mislav/bump-homebrew-formula-action@v1

				        with:

				          formula-name: meilisearch

				        env:

				          COMMITTER_TOKEN: ${{ secrets.HOMEBREW_COMMITTER_TOKEN }}

									
										105

.github/workflows/publish-docker-images.yml
									
										vendored
									
										Normal file
									
												View File
												
				@@ -0,0 +1,105 @@

				name: Publish images to Docker Hub

				on:

				  push:

				    # Will run for every tag pushed except `latest`

				    # When the `latest` git tag is created with this [CI](../latest-git-tag.yml)

				    # we don't need to create a Docker `latest` image again.

				    # The `latest` Docker image push is already done in this CI when releasing a stable version of Meilisearch.

				    tags-ignore:

				      - latest

				  # Both `schedule` and `workflow_dispatch` build the nightly tag

				  schedule:

				    - cron: '0 23 * * *' # Every day at 11:00pm

				  workflow_dispatch:

				jobs:

				  docker:

				    runs-on: docker

				    steps:

				      - uses: actions/checkout@v3

				      # If we are running a cron or manual job ('schedule' or 'workflow_dispatch' event), it means we are publishing the `nightly` tag, so not considered stable.

				      # If we have pushed a tag, and the tag has the v<nmumber>.<number>.<number> format, it means we are publishing an official release, so considered stable.

				      # In this situation, we need to set `output.stable` to create/update the following tags (additionally to the `vX.Y.Z` Docker tag):

				      # - a `vX.Y` (without patch version) Docker tag

				      # - a `latest` Docker tag

				      # For any other tag pushed, this is not considered stable.

				      - name: Define if stable and latest release

				        id: check-tag-format

				        env:

				          # To avoid request limit with the .github/scripts/is-latest-release.sh script

				          GITHUB_PATH: ${{ secrets.MEILI_BOT_GH_PAT }}

				        run: |

				          escaped_tag=$(printf "%q" ${{ github.ref_name }})

				          echo "latest=false" >> $GITHUB_OUTPUT

				          if [[ ${{ github.event_name }} != 'push' ]]; then

				            echo "stable=false" >> $GITHUB_OUTPUT

				          elif [[ $escaped_tag =~ ^v[0-9]+\.[0-9]+\.[0-9]+$ ]]; then

				            echo "stable=true" >> $GITHUB_OUTPUT

				            echo "latest=$(sh .github/scripts/is-latest-release.sh)" >> $GITHUB_OUTPUT

				          else

				            echo "stable=false" >> $GITHUB_OUTPUT

				          fi

				      # Check only the validity of the tag for stable releases (not for pre-releases or other tags)

				      - name: Check release validity

				        if: steps.check-tag-format.outputs.stable == 'true'

				        run: bash .github/scripts/check-release.sh

				      - name: Set build-args for Docker buildx

				        id: build-metadata

				        run: |

				          # Extract commit date

				          commit_date=$(git show -s --format=%cd --date=iso-strict ${{ github.sha }})

				          echo "date=$commit_date" >> $GITHUB_OUTPUT

				      - name: Set up QEMU

				        uses: docker/setup-qemu-action@v3

				      - name: Set up Docker Buildx

				        uses: docker/setup-buildx-action@v3

				      - name: Login to Docker Hub

				        uses: docker/login-action@v3

				        with:

				          username: ${{ secrets.DOCKERHUB_USERNAME }}

				          password: ${{ secrets.DOCKERHUB_TOKEN }}

				      - name: Docker meta

				        id: meta

				        uses: docker/metadata-action@v5

				        with:

				          images: getmeili/meilisearch

				          # Prevent `latest` to be updated for each new tag pushed.

				          # We need latest and `vX.Y` tags to only be pushed for the stable Meilisearch releases.

				          flavor: latest=false

				          tags: |

				            type=ref,event=tag

				            type=raw,value=nightly,enable=${{ github.event_name != 'push' }}

				            type=semver,pattern=v{{major}}.{{minor}},enable=${{ steps.check-tag-format.outputs.stable == 'true' }}

				            type=raw,value=latest,enable=${{ steps.check-tag-format.outputs.stable == 'true' && steps.check-tag-format.outputs.latest == 'true' }}

				      - name: Build and push

				        uses: docker/build-push-action@v5

				        with:

				          push: true

				          platforms: linux/amd64,linux/arm64

				          tags: ${{ steps.meta.outputs.tags }}

				          build-args: |

				            COMMIT_SHA=${{ github.sha }}

				            COMMIT_DATE=${{ steps.build-metadata.outputs.date }}

				            GIT_TAG=${{ github.ref_name }}

				      # /!\ Don't touch this without checking with Cloud team

				      - name: Send CI information to Cloud team

				        # Do not send if nightly build (i.e. 'schedule' or 'workflow_dispatch' event)

				        if: github.event_name == 'push'

				        uses: peter-evans/repository-dispatch@v3

				        with:

				          token: ${{ secrets.MEILI_BOT_GH_PAT }}

				          repository: meilisearch/meilisearch-cloud

				          event-type: cloud-docker-build

				          client-payload: '{ "meilisearch_version": "${{ github.ref_name }}", "stable": "${{ steps.check-tag-format.outputs.stable }}" }'

									
										30

.github/workflows/publish-docker-latest.yml
									
										vendored
									
												View File
											
				@@ -1,30 +0,0 @@

				---

				on:

				  release:

				    types: [released]

				name: Publish latest image to Docker Hub

				jobs:

				  docker-latest:

				    runs-on: docker

				    steps:

				      - name: Set up QEMU

				        uses: docker/setup-qemu-action@v1

				      - name: Set up Docker Buildx

				        uses: docker/setup-buildx-action@v1

				      - name: Login to DockerHub

				        uses: docker/login-action@v1

				        with:

				          username: ${{ secrets.DOCKER_USERNAME }}

				          password: ${{ secrets.DOCKER_PASSWORD }}

				      - name: Build and push

				        id: docker_build

				        uses: docker/build-push-action@v2

				        with:

				          push: true

				          platforms: linux/amd64,linux/arm64

				          tags: getmeili/meilisearch:latest

									
										39

.github/workflows/publish-docker-tag.yml
									
										vendored
									
												View File
											
				@@ -1,39 +0,0 @@

				---

				on:

				  push:

				    tags:

				      - '*'

				name: Publish tagged image to Docker Hub

				jobs:

				  docker-tag:

				    runs-on: docker

				    steps:

				      - name: Set up QEMU

				        uses: docker/setup-qemu-action@v1

				      - name: Set up Docker Buildx

				        uses: docker/setup-buildx-action@v1

				      - name: Login to DockerHub

				        uses: docker/login-action@v1

				        with:

				          username: ${{ secrets.DOCKER_USERNAME }}

				          password: ${{ secrets.DOCKER_PASSWORD }}

				      - name: Docker meta

				        id: meta

				        uses: docker/metadata-action@v3

				        with:

				          images: getmeili/meilisearch

				          flavor: latest=false

				          tags: type=ref,event=tag

				      - name: Build and push

				        id: docker_build

				        uses: docker/build-push-action@v2

				        with:

				          push: true

				          platforms: linux/amd64,linux/arm64

				          tags: ${{ steps.meta.outputs.tags }}

									
										91

.github/workflows/rust.yml
									
										vendored
									
												View File
											
				@@ -1,91 +0,0 @@

				name: Rust

				on:

				  workflow_dispatch:

				  pull_request:

				  push:

				    # trying and staging branches are for Bors config

				    branches:

				      - trying

				      - staging

				env:

				  CARGO_TERM_COLOR: always

				  RUST_BACKTRACE: 1

				jobs:

				  tests:

				    name: Tests on ${{ matrix.os }}

				    runs-on: ${{ matrix.os }}

				    strategy:

				      fail-fast: false

				      matrix:

				        os: [ubuntu-18.04, macos-latest, windows-latest]

				    steps:

				    - uses: actions/checkout@v2

				    - name: Cache dependencies

				      uses: Swatinem/rust-cache@v1.3.0

				    - name: Run cargo check without any default features

				      uses: actions-rs/cargo@v1

				      with:

				        command: build

				        args: --locked --release --no-default-features

				    - name: Run cargo test

				      uses: actions-rs/cargo@v1

				      with:

				        command: test

				        args: --locked --release

				  # We run tests in debug also, to make sure that the debug_assertions are hit

				  test-debug:

				    name: Run tests in debug

				    runs-on: ubuntu-18.04

				    steps:

				      - uses: actions/checkout@v2

				      - uses: actions-rs/toolchain@v1

				        with:

				          profile: minimal

				          toolchain: stable

				          override: true

				      - name: Cache dependencies

				        uses: Swatinem/rust-cache@v1.3.0

				      - name: Run tests in debug

				        uses: actions-rs/cargo@v1

				        with:

				          command: test

				          args: --locked

				  clippy:

				    name: Run Clippy

				    runs-on: ubuntu-18.04

				    steps:

				      - uses: actions/checkout@v2

				      - uses: actions-rs/toolchain@v1

				        with:

				          profile: minimal

				          toolchain: stable

				          override: true

				          components: clippy

				      - name: Cache dependencies

				        uses: Swatinem/rust-cache@v1.3.0

				      - name: Run cargo clippy

				        uses: actions-rs/cargo@v1

				        with:

				          command: clippy

				          args: --all-targets -- --deny warnings

				  fmt:

				    name: Run Rustfmt

				    runs-on: ubuntu-18.04

				    steps:

				      - uses: actions/checkout@v2

				      - uses: actions-rs/toolchain@v1

				        with:

				          profile: minimal

				          toolchain: nightly

				          override: true

				          components: rustfmt

				      - name: Cache dependencies

				        uses: Swatinem/rust-cache@v1.3.0

				      - name: Run cargo fmt

				        run: cargo fmt --all -- --check

									
										386

.github/workflows/sdks-tests.yml
									
										vendored
									
										Normal file
									
												View File
												
				@@ -0,0 +1,386 @@

				# If any test fails, the engine team should ensure the "breaking" changes are expected and contact the integration team

				name: SDKs tests

				on:

				  workflow_dispatch:

				    inputs:

				      docker_image:

				        description: 'The Meilisearch Docker image used'

				        required: false

				        default: nightly

				  schedule:

				    - cron: "0 6 * * MON" # Every Monday at 6:00AM

				env:

				  MEILI_MASTER_KEY: 'masterKey'

				  MEILI_NO_ANALYTICS: 'true'

				  DISABLE_COVERAGE: 'true'

				jobs:

				  define-docker-image:

				    runs-on: ubuntu-latest

				    outputs:

				      docker-image: ${{ steps.define-image.outputs.docker-image }}

				    steps:

				      - uses: actions/checkout@v4

				      - name: Define the Docker image we need to use

				        id: define-image

				        run: |

				          event=${{ github.event_name }}

				          echo "docker-image=nightly" >> $GITHUB_OUTPUT

				          if [[ $event == 'workflow_dispatch' ]]; then

				            echo "docker-image=${{ github.event.inputs.docker_image }}" >> $GITHUB_OUTPUT

				          fi

				      - name: Docker image is ${{ steps.define-image.outputs.docker-image }}

				        run: echo "Docker image is ${{ steps.define-image.outputs.docker-image }}"

				##########

				## SDKs ##

				##########

				  meilisearch-dotnet-tests:

				    needs: define-docker-image

				    name: .NET SDK tests

				    runs-on: ubuntu-latest

				    env:

				      MEILISEARCH_VERSION: ${{ needs.define-docker-image.outputs.docker-image }}

				    steps:

				      - uses: actions/checkout@v4

				        with:

				          repository: meilisearch/meilisearch-dotnet

				      - name: Setup .NET Core

				        uses: actions/setup-dotnet@v4

				        with:

				          dotnet-version: "6.0.x"

				      - name: Install dependencies

				        run: dotnet restore

				      - name: Build

				        run: dotnet build --configuration Release --no-restore

				      - name: Meilisearch (latest version) setup with Docker

				        run: docker compose up -d

				      - name: Run tests

				        run: dotnet test --no-restore --verbosity normal

				  meilisearch-dart-tests:

				    needs: define-docker-image

				    name: Dart SDK tests

				    runs-on: ubuntu-latest

				    services:

				      meilisearch:

				        image: getmeili/meilisearch:${{ needs.define-docker-image.outputs.docker-image }}

				        env:

				          MEILI_MASTER_KEY: ${{ env.MEILI_MASTER_KEY }}

				          MEILI_NO_ANALYTICS: ${{ env.MEILI_NO_ANALYTICS }}

				        ports:

				          - '7700:7700'

				    steps:

				      - uses: actions/checkout@v4

				        with:

				          repository: meilisearch/meilisearch-dart

				      - uses: dart-lang/setup-dart@v1

				        with:

				          sdk: 'latest'

				      - name: Install dependencies

				        run: dart pub get

				      - name: Run integration tests

				        run: dart test --concurrency=4

				  meilisearch-go-tests:

				    needs: define-docker-image

				    name: Go SDK tests

				    runs-on: ubuntu-latest

				    services:

				      meilisearch:

				        image: getmeili/meilisearch:${{ needs.define-docker-image.outputs.docker-image }}

				        env:

				          MEILI_MASTER_KEY: ${{ env.MEILI_MASTER_KEY }}

				          MEILI_NO_ANALYTICS: ${{ env.MEILI_NO_ANALYTICS }}

				        ports:

				          - '7700:7700'

				    steps:

				      - name: Set up Go

				        uses: actions/setup-go@v5

				        with:

				          go-version: stable

				      - uses: actions/checkout@v4

				        with:

				          repository: meilisearch/meilisearch-go

				      - name: Get dependencies

				        run: |

				          go get -v -t -d ./...

				          if [ -f Gopkg.toml ]; then

				            curl https://raw.githubusercontent.com/golang/dep/master/install.sh | sh

				            dep ensure

				          fi

				      - name: Run integration tests

				        run: go test -v ./...

				  meilisearch-java-tests:

				    needs: define-docker-image

				    name: Java SDK tests

				    runs-on: ubuntu-latest

				    services:

				      meilisearch:

				        image: getmeili/meilisearch:${{ needs.define-docker-image.outputs.docker-image }}

				        env:

				          MEILI_MASTER_KEY: ${{ env.MEILI_MASTER_KEY }}

				          MEILI_NO_ANALYTICS: ${{ env.MEILI_NO_ANALYTICS }}

				        ports:

				          - '7700:7700'

				    steps:

				      - uses: actions/checkout@v4

				        with:

				          repository: meilisearch/meilisearch-java

				      - name: Set up Java

				        uses: actions/setup-java@v4

				        with:

				          java-version: 8

				          distribution: 'zulu'

				          cache: gradle

				      - name: Grant execute permission for gradlew

				        run: chmod +x gradlew

				      - name: Build and run unit and integration tests

				        run: ./gradlew build integrationTest

				  meilisearch-js-tests:

				    needs: define-docker-image

				    name: JS SDK tests

				    runs-on: ubuntu-latest

				    services:

				      meilisearch:

				        image: getmeili/meilisearch:${{ needs.define-docker-image.outputs.docker-image }}

				        env:

				          MEILI_MASTER_KEY: ${{ env.MEILI_MASTER_KEY }}

				          MEILI_NO_ANALYTICS: ${{ env.MEILI_NO_ANALYTICS }}

				        ports:

				          - '7700:7700'

				    steps:

				      - uses: actions/checkout@v4

				        with:

				          repository: meilisearch/meilisearch-js

				      - name: Setup node

				        uses: actions/setup-node@v4

				        with:

				          cache: 'yarn'

				      - name: Install dependencies

				        run: yarn --dev

				      - name: Run tests

				        run: yarn test

				      - name: Build project

				        run: yarn build

				      - name: Run ESM env

				        run: yarn test:env:esm

				      - name: Run Node.js env

				        run: yarn test:env:nodejs

				      - name: Run node typescript env

				        run: yarn test:env:node-ts

				      - name: Run Browser env

				        run: yarn test:env:browser

				  meilisearch-php-tests:

				    needs: define-docker-image

				    name: PHP SDK tests

				    runs-on: ubuntu-latest

				    services:

				      meilisearch:

				        image: getmeili/meilisearch:${{ needs.define-docker-image.outputs.docker-image }}

				        env:

				          MEILI_MASTER_KEY: ${{ env.MEILI_MASTER_KEY }}

				          MEILI_NO_ANALYTICS: ${{ env.MEILI_NO_ANALYTICS }}

				        ports:

				          - '7700:7700'

				    steps:

				      - uses: actions/checkout@v4

				        with:

				          repository: meilisearch/meilisearch-php

				      - name: Install PHP

				        uses: shivammathur/setup-php@v2

				      - name: Validate composer.json and composer.lock

				        run: composer validate

				      - name: Install dependencies

				        run: |

				          composer remove --dev friendsofphp/php-cs-fixer --no-update --no-interaction

				          composer update --prefer-dist --no-progress

				      - name: Run test suite - default HTTP client (Guzzle 7)

				        run: |

				          sh scripts/tests.sh

				          composer remove --dev guzzlehttp/guzzle http-interop/http-factory-guzzle

				  meilisearch-python-tests:

				    needs: define-docker-image

				    name: Python SDK tests

				    runs-on: ubuntu-latest

				    services:

				      meilisearch:

				        image: getmeili/meilisearch:${{ needs.define-docker-image.outputs.docker-image }}

				        env:

				          MEILI_MASTER_KEY: ${{ env.MEILI_MASTER_KEY }}

				          MEILI_NO_ANALYTICS: ${{ env.MEILI_NO_ANALYTICS }}

				        ports:

				          - '7700:7700'

				    steps:

				      - uses: actions/checkout@v4

				        with:

				          repository: meilisearch/meilisearch-python

				      - name: Set up Python

				        uses: actions/setup-python@v5

				      - name: Install pipenv

				        uses: dschep/install-pipenv-action@v1

				      - name: Install dependencies

				        run: pipenv install --dev --python=${{ matrix.python-version }}

				      - name: Test with pytest

				        run: pipenv run pytest

				  meilisearch-ruby-tests:

				    needs: define-docker-image

				    name: Ruby SDK tests

				    runs-on: ubuntu-latest

				    services:

				      meilisearch:

				        image: getmeili/meilisearch:${{ needs.define-docker-image.outputs.docker-image }}

				        env:

				          MEILI_MASTER_KEY: ${{ env.MEILI_MASTER_KEY }}

				          MEILI_NO_ANALYTICS: ${{ env.MEILI_NO_ANALYTICS }}

				        ports:

				          - '7700:7700'

				    steps:

				      - uses: actions/checkout@v4

				        with:

				          repository: meilisearch/meilisearch-ruby

				      - name: Set up Ruby 3

				        uses: ruby/setup-ruby@v1

				        with:

				          ruby-version: 3

				      - name: Install ruby dependencies

				        run: bundle install --with test

				      - name: Run test suite

				        run: bundle exec rspec

				  meilisearch-rust-tests:

				    needs: define-docker-image

				    name: Rust SDK tests

				    runs-on: ubuntu-latest

				    services:

				      meilisearch:

				        image: getmeili/meilisearch:${{ needs.define-docker-image.outputs.docker-image }}

				        env:

				          MEILI_MASTER_KEY: ${{ env.MEILI_MASTER_KEY }}

				          MEILI_NO_ANALYTICS: ${{ env.MEILI_NO_ANALYTICS }}

				        ports:

				          - '7700:7700'

				    steps:

				      - uses: actions/checkout@v4

				        with:

				          repository: meilisearch/meilisearch-rust

				      - name: Build

				        run: cargo build --verbose

				      - name: Run tests

				        run: cargo test --verbose

				  meilisearch-swift-tests:

				    needs: define-docker-image

				    name: Swift SDK tests

				    runs-on: ubuntu-latest

				    services:

				      meilisearch:

				        image: getmeili/meilisearch:${{ needs.define-docker-image.outputs.docker-image }}

				        env:

				          MEILI_MASTER_KEY: ${{ env.MEILI_MASTER_KEY }}

				          MEILI_NO_ANALYTICS: ${{ env.MEILI_NO_ANALYTICS }}

				        ports:

				          - '7700:7700'

				    steps:

				      - uses: actions/checkout@v4

				        with:

				          repository: meilisearch/meilisearch-swift

				      - name: Run tests

				        run: swift test

				########################

				## FRONT-END PLUGINS ##

				########################

				  meilisearch-js-plugins-tests:

				    needs: define-docker-image

				    name: meilisearch-js-plugins tests

				    runs-on: ubuntu-latest

				    services:

				      meilisearch:

				        image: getmeili/meilisearch:${{ needs.define-docker-image.outputs.docker-image }}

				        env:

				          MEILI_MASTER_KEY: ${{ env.MEILI_MASTER_KEY }}

				          MEILI_NO_ANALYTICS: ${{ env.MEILI_NO_ANALYTICS }}

				        ports:

				          - '7700:7700'

				    steps:

				      - uses: actions/checkout@v4

				        with:

				          repository: meilisearch/meilisearch-js-plugins

				      - name: Setup node

				        uses: actions/setup-node@v4

				        with:

				          cache: yarn

				      - name: Install dependencies

				        run: yarn install

				      - name: Run tests

				        run: yarn test

				      - name: Build all the playgrounds and the packages

				        run: yarn build

				########################

				## BACK-END PLUGINS ###

				########################

				  meilisearch-rails-tests:

				    needs: define-docker-image

				    name: meilisearch-rails tests

				    runs-on: ubuntu-latest

				    services:

				      meilisearch:

				        image: getmeili/meilisearch:${{ needs.define-docker-image.outputs.docker-image }}

				        env:

				          MEILI_MASTER_KEY: ${{ env.MEILI_MASTER_KEY }}

				          MEILI_NO_ANALYTICS: ${{ env.MEILI_NO_ANALYTICS }}

				        ports:

				          - '7700:7700'

				    steps:

				      - uses: actions/checkout@v4

				        with:

				          repository: meilisearch/meilisearch-rails

				      - name: Set up Ruby 3

				        uses: ruby/setup-ruby@v1

				        with:

				          ruby-version: 3

				          bundler-cache: true

				      - name: Run tests

				        run: bundle exec rspec

				  meilisearch-symfony-tests:

				    needs: define-docker-image

				    name: meilisearch-symfony tests

				    runs-on: ubuntu-latest

				    services:

				      meilisearch:

				        image: getmeili/meilisearch:${{ needs.define-docker-image.outputs.docker-image }}

				        env:

				          MEILI_MASTER_KEY: ${{ env.MEILI_MASTER_KEY }}

				          MEILI_NO_ANALYTICS: ${{ env.MEILI_NO_ANALYTICS }}

				        ports:

				          - '7700:7700'

				    steps:

				      - uses: actions/checkout@v4

				        with:

				          repository: meilisearch/meilisearch-symfony

				      - name: Install PHP

				        uses: shivammathur/setup-php@v2

				        with:

				          tools: composer:v2, flex

				      - name: Validate composer.json and composer.lock

				        run: composer validate

				      - name: Install dependencies

				        run: composer install --prefer-dist --no-progress --quiet

				      - name: Remove doctrine/annotations

				        run: composer remove --dev doctrine/annotations

				      - name: Run test suite

				        run: composer test:unit

									
										190

.github/workflows/test-suite.yml
									
										vendored
									
										Normal file
									
												View File
												
				@@ -0,0 +1,190 @@

				name: Test suite

				on:

				  workflow_dispatch:

				  schedule:

				    # Everyday at 5:00am

				    - cron: '0 5 * * *'

				  pull_request:

				  push:

				    # trying and staging branches are for Bors config

				    branches:

				      - trying

				      - staging

				env:

				  CARGO_TERM_COLOR: always

				  RUST_BACKTRACE: 1

				  RUSTFLAGS: "-D warnings"

				jobs:

				  test-linux:

				    name: Tests on ubuntu-18.04

				    runs-on: ubuntu-latest

				    container:

				      # Use ubuntu-18.04 to compile with glibc 2.27, which are the production expectations

				      image: ubuntu:18.04

				    steps:

				      - uses: actions/checkout@v3

				      - name: Install needed dependencies

				        run: |

				          apt-get update && apt-get install -y curl

				          apt-get install build-essential -y

				      - name: Setup test with Rust stable

				        uses: actions-rs/toolchain@v1

				        with:

				          toolchain: stable

				          override: true

				      - name: Cache dependencies

				        uses: Swatinem/rust-cache@v2.7.1

				      - name: Run cargo check without any default features

				        uses: actions-rs/cargo@v1

				        with:

				          command: build

				          args: --locked --release --no-default-features --all

				      - name: Run cargo test

				        uses: actions-rs/cargo@v1

				        with:

				          command: test

				          args: --locked --release --all

				  test-others:

				    name: Tests on ${{ matrix.os }}

				    runs-on: ${{ matrix.os }}

				    strategy:

				      fail-fast: false

				      matrix:

				        os: [macos-12, windows-2022]

				    steps:

				      - uses: actions/checkout@v3

				      - name: Cache dependencies

				        uses: Swatinem/rust-cache@v2.7.1

				      - uses: actions-rs/toolchain@v1

				        with:

				          toolchain: stable

				          override: true

				      - name: Run cargo check without any default features

				        uses: actions-rs/cargo@v1

				        with:

				          command: build

				          args: --locked --release --no-default-features --all

				      - name: Run cargo test

				        uses: actions-rs/cargo@v1

				        with:

				          command: test

				          args: --locked --release --all

				  test-all-features:

				    name: Tests almost all features

				    runs-on: ubuntu-latest

				    container:

				      # Use ubuntu-18.04 to compile with glibc 2.27, which are the production expectations

				      image: ubuntu:18.04

				    if: github.event_name == 'schedule' || github.event_name == 'workflow_dispatch'

				    steps:

				      - uses: actions/checkout@v3

				      - name: Install needed dependencies

				        run: |

				          apt-get update

				          apt-get install --assume-yes build-essential curl

				      - uses: actions-rs/toolchain@v1

				        with:

				          toolchain: stable

				          override: true

				      - name: Run cargo build with almost all features

				        run: |

				          cargo build --workspace --locked --release --features "$(cargo xtask list-features --exclude-feature cuda)"

				      - name: Run cargo test with almost all features

				        run: |

				          cargo test --workspace --locked --release --features "$(cargo xtask list-features --exclude-feature cuda)"

				  test-disabled-tokenization:

				    name: Test disabled tokenization

				    runs-on: ubuntu-latest

				    container:

				      image: ubuntu:18.04

				    if: github.event_name == 'schedule' || github.event_name == 'workflow_dispatch'

				    steps:

				      - uses: actions/checkout@v3

				      - name: Install needed dependencies

				        run: |

				          apt-get update

				          apt-get install --assume-yes build-essential curl

				      - uses: actions-rs/toolchain@v1

				        with:

				          toolchain: stable

				          override: true

				      - name: Run cargo tree without default features and check lindera is not present

				        run: |

				          if cargo tree -f '{p} {f}' -e normal --no-default-features | grep -vqz lindera; then

				            echo "lindera has been found in the sources and it shouldn't"

				            exit 1

				          fi

				      - name: Run cargo tree with default features and check lindera is pressent

				        run: |

				          cargo tree -f '{p} {f}' -e normal | grep lindera -qz

				  # We run tests in debug also, to make sure that the debug_assertions are hit

				  test-debug:

				    name: Run tests in debug

				    runs-on: ubuntu-latest

				    container:

				      # Use ubuntu-18.04 to compile with glibc 2.27, which are the production expectations

				      image: ubuntu:18.04

				    steps:

				      - uses: actions/checkout@v3

				      - name: Install needed dependencies

				        run: |

				          apt-get update && apt-get install -y curl

				          apt-get install build-essential -y

				      - uses: actions-rs/toolchain@v1

				        with:

				          toolchain: stable

				          override: true

				      - name: Cache dependencies

				        uses: Swatinem/rust-cache@v2.7.1

				      - name: Run tests in debug

				        uses: actions-rs/cargo@v1

				        with:

				          command: test

				          args: --locked --all

				  clippy:

				    name: Run Clippy

				    runs-on: ubuntu-latest

				    steps:

				      - uses: actions/checkout@v3

				      - uses: actions-rs/toolchain@v1

				        with:

				          profile: minimal

				          toolchain: 1.75.0

				          override: true

				          components: clippy

				      - name: Cache dependencies

				        uses: Swatinem/rust-cache@v2.7.1

				      - name: Run cargo clippy

				        uses: actions-rs/cargo@v1

				        with:

				          command: clippy

				          args: --all-targets -- --deny warnings

				  fmt:

				    name: Run Rustfmt

				    runs-on: ubuntu-latest

				    steps:

				      - uses: actions/checkout@v3

				      - uses: actions-rs/toolchain@v1

				        with:

				          profile: minimal

				          toolchain: nightly

				          override: true

				          components: rustfmt

				      - name: Cache dependencies

				        uses: Swatinem/rust-cache@v2.7.1

				      - name: Run cargo fmt

				        # Since we never ran the `build.rs` script in the benchmark directory we are missing one auto-generated import file.

				        # Since we want to trigger (and fail) this action as fast as possible, instead of building the benchmark crate

				        # we are going to create an empty file where rustfmt expects it.

				        run: |

				          echo -ne "\n" > benchmarks/benches/datasets_paths.rs

				          cargo fmt --all -- --check

									
										47

.github/workflows/update-cargo-toml-version.yml
									
										vendored
									
										Normal file
									
												View File
												
				@@ -0,0 +1,47 @@

				name: Update Meilisearch version in Cargo.toml

				on:

				  workflow_dispatch:

				    inputs:

				      new_version:

				        description: 'The new version (vX.Y.Z)'

				        required: true

				env:

				  NEW_VERSION: ${{ github.event.inputs.new_version }}

				  NEW_BRANCH: update-version-${{ github.event.inputs.new_version }}

				  GH_TOKEN: ${{ secrets.MEILI_BOT_GH_PAT }}

				jobs:

				  update-version-cargo-toml:

				    name: Update version in Cargo.toml

				    runs-on: ubuntu-latest

				    steps:

				      - uses: actions/checkout@v3

				      - uses: actions-rs/toolchain@v1

				        with:

				          profile: minimal

				          toolchain: stable

				          override: true

				      - name: Install sd

				        run: cargo install sd

				      - name: Update Cargo.toml file

				        run: |

				          raw_new_version=$(echo $NEW_VERSION | cut -d 'v' -f 2)

				          new_string="version = \"$raw_new_version\""

				          sd '^version = "\d+.\d+.\w+"$' "$new_string" Cargo.toml

				      - name: Build Meilisearch to update Cargo.lock

				        run: cargo build

				      - name: Commit and push the changes to the ${{ env.NEW_BRANCH }} branch

				        uses: EndBug/add-and-commit@v9

				        with:

				          message: "Update version for the next release (${{ env.NEW_VERSION }}) in Cargo.toml"

				          new_branch: ${{ env.NEW_BRANCH }}

				      - name: Create the PR pointing to ${{ github.ref_name }}

				        run: |

				          gh pr create \

				            --title "Update version for the next release ($NEW_VERSION) in Cargo.toml" \

				            --body '⚠️ This PR is automatically generated. Check the new version is the expected one and Cargo.lock has been updated before merging.' \

				            --label 'skip changelog' \

				            --milestone $NEW_VERSION \

				            --base $GITHUB_REF_NAME

13

.gitignore vendored

View File

@@ -1,3 +1,5 @@
 .idea/
 .vscode/
 /target
 **/*.csv
 **/*.json_lines
@@ -7,3 +9,14 @@
 /data.ms
 /snapshots
 /dumps
 /bench
 /_xtask_benchmark.ms
 # Snapshots
 ## ... large
 *.full.snap
 ##  ... unreviewed
 *.snap.new
 # Fuzzcheck data for the facet indexing fuzz test
 milli/fuzz/update::facet::incremental::fuzz::fuzz/

									
										5

.rustfmt.toml
									
										Normal file
									
												View File
												
				@@ -0,0 +1,5 @@

				unstable_features = true

				use_small_heuristics = "max"

				imports_granularity = "Module"

				group_imports = "StdExternalCrate"

									
										72

CONTRIBUTING.md
									
												View File
												
				@@ -4,17 +4,23 @@ First, thank you for contributing to Meilisearch! The goal of this document is t

				Remember that there are many ways to contribute other than writing code: writing [tutorials or blog posts](https://github.com/meilisearch/awesome-meilisearch), improving [the documentation](https://github.com/meilisearch/documentation), submitting [bug reports](https://github.com/meilisearch/meilisearch/issues/new?assignees=&labels=&template=bug_report.md&title=) and [feature requests](https://github.com/meilisearch/product/discussions/categories/feedback-feature-proposal)...

				The code in this repository is only concerned with managing multiple indexes, handling the update store, and exposing an HTTP API. Search and indexation are the domain of our core engine, [`milli`](https://github.com/meilisearch/milli), while tokenization is handled by [our `charabia` library](https://github.com/meilisearch/charabia/).

				If Meilisearch does not offer optimized support for your language, please consider contributing to `charabia` by following the [CONTRIBUTING.md file](https://github.com/meilisearch/charabia/blob/main/CONTRIBUTING.md) and integrating your intended normalizer/segmenter.

				## Table of Contents

				- [Assumptions](#assumptions)

				- [How to Contribute](#how-to-contribute)

				- [Development Workflow](#development-workflow)

				- [Git Guidelines](#git-guidelines)

				- [Release Process (for internal team only)](#release-process-for-internal-team-only)

				## Assumptions

				1. **You're familiar with [Github](https://github.com) and the [Pull Requests](https://help.github.com/en/github/collaborating-with-issues-and-pull-requests/about-pull-requests)(PR) workflow.**

				2. **You've read the Meilisearch [documentation](https://docs.meilisearch.com).**

				3. **You know about the [Meilisearch community](https://docs.meilisearch.com/learn/what_is_meilisearch/contact.html).

				1. **You're familiar with [GitHub](https://github.com) and the [Pull Requests (PR)](https://help.github.com/en/github/collaborating-with-issues-and-pull-requests/about-pull-requests) workflow.**

				2. **You've read the Meilisearch [documentation](https://www.meilisearch.com/docs).**

				3. **You know about the [Meilisearch community on Discord](https://discord.meilisearch.com).

				   Please use this for help.**

				## How to Contribute

				@@ -22,7 +28,7 @@ Remember that there are many ways to contribute other than writing code: writing

				1. Ensure your change has an issue! Find an

				   [existing issue](https://github.com/meilisearch/meilisearch/issues/) or [open a new issue](https://github.com/meilisearch/meilisearch/issues/new).

				   * This is where you can get a feel if the change will be accepted or not.

				2. Once approved, [fork the Meilisearch repository](https://help.github.com/en/github/getting-started-with-github/fork-a-repo) in your own Github account.

				2. Once approved, [fork the Meilisearch repository](https://help.github.com/en/github/getting-started-with-github/fork-a-repo) in your own GitHub account.

				3. [Create a new Git branch](https://help.github.com/en/github/collaborating-with-issues-and-pull-requests/creating-and-deleting-branches-within-your-repository)

				4. Review the [Development Workflow](#development-workflow) section that describes the steps to maintain the repository.

				5. Make your changes on your branch.

				@@ -44,12 +50,37 @@ We recommend using the `--release` flag to test the full performance of Meilisea

				cargo test

				```

				This command will be triggered to each PR as a requirement for merging it.

				#### Snapshot-based tests

				We are using [insta](https://insta.rs) to perform snapshot-based testing.

				We recommend using the insta tooling (such as `cargo-insta`) to update the snapshots if they change following a PR.

				New tests should use insta where possible rather than manual `assert` statements.

				Furthermore, we provide some macros on top of insta, notably a way to use snapshot hashes instead of inline snapshots, saving a lot of space in the repository.

				To effectively debug snapshot-based hashes, we recommend you export the `MEILI_TEST_FULL_SNAPS` environment variable so that snapshot are fully created locally:

				```

				export MEILI_TEST_FULL_SNAPS=true # add this to your .bashrc, .zshrc, ...

				```

				#### Test troubleshooting

				If you get a "Too many open files" error you might want to increase the open file limit using this command:

				```bash

				ulimit -Sn 3000

				```

				#### Build tools

				Meilisearch follows the [cargo xtask](https://github.com/matklad/cargo-xtask) workflow to provide some build tools.

				Run `cargo xtask --help` from the root of the repository to find out what is available.

				## Git Guidelines

				### Git Branches

				@@ -68,7 +99,7 @@ As minimal requirements, your commit message should:

				We don't follow any other convention, but if you want to use one, we recommend [the Chris Beams one](https://chris.beams.io/posts/git-commit/).

				### Github Pull Requests

				### GitHub Pull Requests

				Some notes on GitHub PRs:

				@@ -78,6 +109,37 @@ Some notes on GitHub PRs:

				  The draft PRs are recommended when you want to show that you are working on something and make your work visible.

				- The branch related to the PR must be **up-to-date with `main`** before merging. Fortunately, this project uses [Bors](https://github.com/bors-ng/bors-ng) to automatically enforce this requirement without the PR author having to rebase manually.

				## Release Process (for internal team only)

				Meilisearch tools follow the [Semantic Versioning Convention](https://semver.org/).

				### Automation to rebase and Merge the PRs

				This project integrates a bot that helps us manage pull requests merging.<br>

				_[Read more about this](https://github.com/meilisearch/integration-guides/blob/main/resources/bors.md)._

				### How to Publish a new Release

				The full Meilisearch release process is described in [this guide](https://github.com/meilisearch/engine-team/blob/main/resources/meilisearch-release.md). Please follow it carefully before doing any release.

				### How to publish a prototype

				Depending on the developed feature, you might need to provide a prototyped version of Meilisearch to make it easier to test by the users.

				This happens in two steps:

				- [Release the prototype](https://github.com/meilisearch/engine-team/blob/main/resources/prototypes.md#how-to-publish-a-prototype)

				- [Communicate about it](https://github.com/meilisearch/engine-team/blob/main/resources/prototypes.md#communication)

				### Release assets

				For each release, the following assets are created:

				- Binaries for different platforms (Linux, MacOS, Windows and ARM architectures) are attached to the GitHub release

				- Binaries are pushed to HomeBrew and APT (not published for RC)

				- Docker tags are created/updated:

				  - `vX.Y.Z`

				  - `vX.Y` (not published for RC)

				  - `latest` (not published for RC)

				<hr>

				Thank you again for reading this through, we can not wait to begin to work with you if you made your way through this contributing guide ❤️

4874

Cargo.lock generated

View File

File diff suppressed because it is too large Load Diff

									
										62

Cargo.toml
									
												View File
												
				@@ -1,9 +1,65 @@

				[workspace]

				resolver = "2"

				members = [

				    "meilisearch-http",

				    "meilisearch-error",

				    "meilisearch-lib",

				    "meilisearch",

				    "meilitool",

				    "meilisearch-types",

				    "meilisearch-auth",

				    "meili-snap",

				    "index-scheduler",

				    "dump",

				    "file-store",

				    "permissive-json-pointer",

				    "milli",

				    "filter-parser",

				    "flatten-serde-json",

				    "json-depth-checker",

				    "benchmarks",

				    "fuzzers",

				    "tracing-trace",

				    "xtask", "build-info",

				]

				[workspace.package]

				version = "1.7.0"

				authors = [

				    "Quentin de Quelen <quentin@dequelen.me>",

				    "Clément Renault <clement@meilisearch.com>",

				]

				description = "Meilisearch HTTP server"

				homepage = "https://meilisearch.com"

				readme = "README.md"

				edition = "2021"

				license = "MIT"

				[profile.release]

				codegen-units = 1

				[profile.dev.package.flate2]

				opt-level = 3

				[profile.dev.package.grenad]

				opt-level = 3

				[profile.dev.package.roaring]

				opt-level = 3

				[profile.dev.package.lindera-ipadic-builder]

				opt-level = 3

				[profile.dev.package.encoding]

				opt-level = 3

				[profile.dev.package.yada]

				opt-level = 3

				[profile.release.package.lindera-ipadic-builder]

				opt-level = 3

				[profile.release.package.encoding]

				opt-level = 3

				[profile.release.package.yada]

				opt-level = 3

				[profile.bench.package.lindera-ipadic-builder]

				opt-level = 3

				[profile.bench.package.encoding]

				opt-level = 3

				[profile.bench.package.yada]

				opt-level = 3

									
										18

Dockerfile
									
												View File
												
				@@ -1,13 +1,14 @@

				# Compile

				FROM    rust:alpine3.14 AS compiler

				FROM    rust:1.75.0-alpine3.18 AS compiler

				RUN     apk add -q --update-cache --no-cache build-base openssl-dev

				WORKDIR /meilisearch

				WORKDIR /

				ARG     COMMIT_SHA

				ARG     COMMIT_DATE

				ENV     COMMIT_SHA=${COMMIT_SHA} COMMIT_DATE=${COMMIT_DATE}

				ARG     GIT_TAG

				ENV     VERGEN_GIT_SHA=${COMMIT_SHA} VERGEN_GIT_COMMIT_TIMESTAMP=${COMMIT_DATE} VERGEN_GIT_DESCRIBE=${GIT_TAG}

				ENV     RUSTFLAGS="-C target-feature=-crt-static"

				COPY    . .

				@@ -16,10 +17,10 @@ RUN     set -eux; \

				        if [ "$apkArch" = "aarch64" ]; then \

				            export JEMALLOC_SYS_WITH_LG_PAGE=16; \

				        fi && \

				        cargo build --release

				        cargo build --release -p meilisearch -p meilitool

				# Run

				FROM    alpine:3.14

				FROM    alpine:3.16

				ENV     MEILI_HTTP_ADDR 0.0.0.0:7700

				ENV     MEILI_SERVER_PROVIDER docker

				@@ -27,9 +28,10 @@ ENV     MEILI_SERVER_PROVIDER docker

				RUN     apk update --quiet \

				        && apk add -q --no-cache libgcc tini curl

				# add meilisearch to the `/bin` so you can run it from anywhere and it's easy

				# to find.

				COPY    --from=compiler /meilisearch/target/release/meilisearch /bin/meilisearch

				# add meilisearch and meilitool to the `/bin` so you can run it from anywhere

				# and it's easy to find.

				COPY    --from=compiler /target/release/meilisearch /bin/meilisearch

				COPY    --from=compiler /target/release/meilitool /bin/meilitool

				# To stay compatible with the older version of the container (pre v0.27.0) we're

				# going to symlink the meilisearch binary in the path to `/meilisearch`

				RUN     ln -s /bin/meilisearch /meilisearch

2

LICENSE

View File

@@ -1,6 +1,6 @@
 MIT License
 Copyright (c) 2019-2022 Meili SAS
 Copyright (c) 2019-2024 Meili SAS
 Permission is hereby granted, free of charge, to any person obtaining a copy
 of this software and associated documentation files (the "Software"), to deal

									
										19

PROFILING.md
									
										Normal file
									
												View File
												
				@@ -0,0 +1,19 @@

				# Profiling Meilisearch

				Search engine technologies are complex pieces of software that require thorough profiling tools. We chose to use [Puffin](https://github.com/EmbarkStudios/puffin), which the Rust gaming industry uses extensively. You can export and import the profiling reports using the top bar's _File_ menu options [in Puffin Viewer](https://github.com/embarkstudios/puffin#ui).

				![An example profiling with Puffin viewer](assets/profiling-example.png)

				## Profiling the Indexing Process

				When you enable [the `exportPuffinReports` experimental feature](https://www.meilisearch.com/docs/learn/experimental/overview) of Meilisearch, Puffin reports with the `.puffin` extension will be automatically exported to disk. When this option is enabled, the engine will automatically create a "frame" whenever it executes the `IndexScheduler::tick` method.

				[Puffin Viewer](https://github.com/EmbarkStudios/puffin/tree/main/puffin_viewer) is used to analyze the reports. Those reports show areas where Meilisearch spent time during indexing.

				Another piece of advice on the Puffin viewer UI interface is to consider the _Merge children with same ID_ option. It can hide the exact actual timings at which events were sent. Please turn it off when you see strange gaps on the Flamegraph. It can help.

				## Profiling the Search Process

				We still need to take the time to profile the search side of the engine with Puffin. It would require time to profile the filtering phase, query parsing, creation, and execution. We could even profile the Actix HTTP server.

				The only issue we see is the framing system. Puffin requires a global frame-based profiling phase, which collides with Meilisearch's ability to accept and answer multiple requests on different threads simultaneously.

									
										225

README.md
									
												View File
												
				@@ -1,205 +1,116 @@

				<p align="center">

				  <img src="assets/logo.svg" alt="Meilisearch" width="200" height="200" />

				  <a href="https://www.meilisearch.com/?utm_campaign=oss&utm_source=github&utm_medium=meilisearch&utm_content=logo#gh-light-mode-only" target="_blank">

				    <img src="assets/meilisearch-logo-light.svg?sanitize=true#gh-light-mode-only">

				  </a>

				  <a href="https://www.meilisearch.com/?utm_campaign=oss&utm_source=github&utm_medium=meilisearch&utm_content=logo#gh-dark-mode-only" target="_blank">

				    <img src="assets/meilisearch-logo-dark.svg?sanitize=true#gh-dark-mode-only">

				  </a>

				</p>

				<h1 align="center">Meilisearch</h1>

				<h4 align="center">

				  <a href="https://www.meilisearch.com">Website</a> |

				  <a href="https://www.meilisearch.com/?utm_campaign=oss&utm_source=github&utm_medium=meilisearch&utm_content=nav">Website</a> |

				  <a href="https://roadmap.meilisearch.com/tabs/1-under-consideration">Roadmap</a> |

				  <a href="https://blog.meilisearch.com">Blog</a> |

				  <a href="https://fr.linkedin.com/company/meilisearch">LinkedIn</a> |

				  <a href="https://twitter.com/meilisearch">Twitter</a> |

				  <a href="https://docs.meilisearch.com">Documentation</a> |

				  <a href="https://docs.meilisearch.com/faq/">FAQ</a>

				  <a href="https://www.meilisearch.com/pricing?utm_campaign=oss&utm_source=github&utm_medium=meilisearch&utm_content=nav">Meilisearch Cloud</a> |

				  <a href="https://blog.meilisearch.com/?utm_campaign=oss&utm_source=github&utm_medium=meilisearch&utm_content=nav">Blog</a> |

				  <a href="https://www.meilisearch.com/docs?utm_campaign=oss&utm_source=github&utm_medium=meilisearch&utm_content=nav">Documentation</a> |

				  <a href="https://www.meilisearch.com/docs/faq?utm_campaign=oss&utm_source=github&utm_medium=meilisearch&utm_content=nav">FAQ</a> |

				  <a href="https://discord.meilisearch.com/?utm_campaign=oss&utm_source=github&utm_medium=meilisearch&utm_content=nav">Discord</a>

				</h4>

				<p align="center">

				  <a href="https://github.com/meilisearch/meilisearch/actions"><img src="https://github.com/meilisearch/meilisearch/workflows/Cargo%20test/badge.svg" alt="Build Status"></a>

				  <a href="https://deps.rs/repo/github/meilisearch/meilisearch"><img src="https://deps.rs/repo/github/meilisearch/meilisearch/status.svg" alt="Dependency status"></a>

				  <a href="https://github.com/meilisearch/meilisearch/blob/main/LICENSE"><img src="https://img.shields.io/badge/license-MIT-informational" alt="License"></a>

				  <a href="https://slack.meilisearch.com"><img src="https://img.shields.io/badge/slack-meilisearch-blue.svg?logo=slack" alt="Slack"></a>

				  <a href="https://github.com/meilisearch/meilisearch/discussions" alt="Discussions"><img src="https://img.shields.io/badge/github-discussions-red" /></a>

				  <a href="https://app.bors.tech/repositories/26457"><img src="https://bors.tech/images/badge_small.svg" alt="Bors enabled"></a>

				  <a href="https://ms-bors.herokuapp.com/repositories/52"><img src="https://bors.tech/images/badge_small.svg" alt="Bors enabled"></a>

				</p>

				<p align="center">⚡ Lightning Fast, Ultra Relevant, and Typo-Tolerant Search Engine 🔍</p>

				<p align="center">⚡ A lightning-fast search engine that fits effortlessly into your apps, websites, and workflow 🔍</p>

				**Meilisearch** is a powerful, fast, open-source, easy to use and deploy search engine. Both searching and indexing are highly customizable. Features such as typo-tolerance, filters, and synonyms are provided out-of-the-box.

				For more information about features go to [our documentation](https://docs.meilisearch.com/).

				Meilisearch helps you shape a delightful search experience in a snap, offering features that work out-of-the-box to speed up your workflow.

				<p align="center">

				  <img src="assets/trumen-fast.gif" alt="Web interface gif" />

				<p align="center" name="demo">

				  <a href="https://where2watch.meilisearch.com/?utm_campaign=oss&utm_source=github&utm_medium=meilisearch&utm_content=demo-gif#gh-light-mode-only" target="_blank">

				    <img src="assets/demo-light.gif#gh-light-mode-only" alt="A bright colored application for finding movies screening near the user">

				  </a>

				  <a href="https://where2watch.meilisearch.com/?utm_campaign=oss&utm_source=github&utm_medium=meilisearch&utm_content=demo-gif#gh-dark-mode-only" target="_blank">

				    <img src="assets/demo-dark.gif#gh-dark-mode-only" alt="A dark colored application for finding movies screening near the user">

				  </a>

				</p>

				🔥 [**Try it!**](https://where2watch.meilisearch.com/?utm_campaign=oss&utm_source=github&utm_medium=meilisearch&utm_content=demo-link) 🔥

				## ✨ Features

				* Search-as-you-type experience (answers < 50 milliseconds)

				* Full-text search

				* Typo tolerant (understands typos and misspelling)

				* Faceted search and filters

				* Supports hanzi (Chinese characters)

				* Supports synonyms

				* Easy to install, deploy, and maintain

				* Whole documents are returned

				* Highly customizable

				* RESTful API

				## Getting started

				- **Search-as-you-type:** find search results in less than 50 milliseconds

				- **[Typo tolerance](https://www.meilisearch.com/docs/learn/configuration/typo_tolerance?utm_campaign=oss&utm_source=github&utm_medium=meilisearch&utm_content=features):** get relevant matches even when queries contain typos and misspellings

				- **[Filtering](https://www.meilisearch.com/docs/learn/fine_tuning_results/filtering?utm_campaign=oss&utm_source=github&utm_medium=meilisearch&utm_content=features) and [faceted search](https://www.meilisearch.com/docs/learn/fine_tuning_results/faceted_search?utm_campaign=oss&utm_source=github&utm_medium=meilisearch&utm_content=features):** enhance your users' search experience with custom filters and build a faceted search interface in a few lines of code

				- **[Sorting](https://www.meilisearch.com/docs/learn/fine_tuning_results/sorting?utm_campaign=oss&utm_source=github&utm_medium=meilisearch&utm_content=features):** sort results based on price, date, or pretty much anything else your users need

				- **[Synonym support](https://www.meilisearch.com/docs/learn/configuration/synonyms?utm_campaign=oss&utm_source=github&utm_medium=meilisearch&utm_content=features):** configure synonyms to include more relevant content in your search results

				- **[Geosearch](https://www.meilisearch.com/docs/learn/fine_tuning_results/geosearch?utm_campaign=oss&utm_source=github&utm_medium=meilisearch&utm_content=features):** filter and sort documents based on geographic data

				- **[Extensive language support](https://www.meilisearch.com/docs/learn/what_is_meilisearch/language?utm_campaign=oss&utm_source=github&utm_medium=meilisearch&utm_content=features):** search datasets in any language, with optimized support for Chinese, Japanese, Hebrew, and languages using the Latin alphabet

				- **[Security management](https://www.meilisearch.com/docs/learn/security/master_api_keys?utm_campaign=oss&utm_source=github&utm_medium=meilisearch&utm_content=features):** control which users can access what data with API keys that allow fine-grained permissions handling

				- **[Multi-Tenancy](https://www.meilisearch.com/docs/learn/security/tenant_tokens?utm_campaign=oss&utm_source=github&utm_medium=meilisearch&utm_content=features):** personalize search results for any number of application tenants

				- **Highly Customizable:** customize Meilisearch to your specific needs or use our out-of-the-box and hassle-free presets

				- **[RESTful API](https://www.meilisearch.com/docs/reference/api/overview?utm_campaign=oss&utm_source=github&utm_medium=meilisearch&utm_content=features):** integrate Meilisearch in your technical stack with our plugins and SDKs

				- **Easy to install, deploy, and maintain**

				### Deploy the Server

				## 📖 Documentation

				#### Homebrew (Mac OS)

				You can consult Meilisearch's documentation at [https://www.meilisearch.com/docs](https://www.meilisearch.com/docs/?utm_campaign=oss&utm_source=github&utm_medium=meilisearch&utm_content=docs).

				```bash

				brew update && brew install meilisearch

				meilisearch

				```

				## 🚀 Getting started

				#### Docker

				For basic instructions on how to set up Meilisearch, add documents to an index, and search for documents, take a look at our [Quick Start](https://www.meilisearch.com/docs/learn/getting_started/quick_start?utm_campaign=oss&utm_source=github&utm_medium=meilisearch&utm_content=get-started) guide.

				```bash

				docker run -p 7700:7700 -v "$(pwd)/data.ms:/data.ms" getmeili/meilisearch

				```

				## ⚡ Supercharge your Meilisearch experience

				#### Announcing a cloud-hosted Meilisearch

				Say goodbye to server deployment and manual updates with [Meilisearch Cloud](https://www.meilisearch.com/cloud?utm_campaign=oss&utm_source=github&utm_medium=meilisearch). No credit card required.

				Join the closed beta by filling out this [form](https://meilisearch.typeform.com/to/FtnzvZfh).

				## 🧰 SDKs & integration tools

				#### Try Meilisearch in our Sandbox

				Install one of our SDKs in your project for seamless integration between Meilisearch and your favorite language or framework!

				Create a Meilisearch instance in [Meilisearch Sandbox](https://sandbox.meilisearch.com/). This instance is free, and will be active for 48 hours.

				Take a look at the complete [Meilisearch integration list](https://www.meilisearch.com/docs/learn/what_is_meilisearch/sdks?utm_campaign=oss&utm_source=github&utm_medium=meilisearch&utm_content=sdks-link).

				#### Run on Digital Ocean

				[![Logos belonging to different languages and frameworks supported by Meilisearch, including React, Ruby on Rails, Go, Rust, and PHP](assets/integrations.png)](https://www.meilisearch.com/docs/learn/what_is_meilisearch/sdks?utm_campaign=oss&utm_source=github&utm_medium=meilisearch&utm_content=sdks-logos)

				[![DigitalOcean Marketplace](assets/do-btn-blue.svg)](https://marketplace.digitalocean.com/apps/meilisearch?action=deploy&refcode=7c67bd97e101)

				## ⚙️ Advanced usage

				#### Deploy on Platform.sh

				Experienced users will want to keep our [API Reference](https://www.meilisearch.com/docs/reference/api/overview?utm_campaign=oss&utm_source=github&utm_medium=meilisearch&utm_content=advanced) close at hand.

				<a href="https://console.platform.sh/projects/create-project?template=https://raw.githubusercontent.com/platformsh/template-builder/master/templates/meilisearch/.platform.template.yaml&utm_content=meilisearch&utm_source=github&utm_medium=button&utm_campaign=deploy_on_platform">

				    <img src="https://platform.sh/images/deploy/lg-blue.svg" alt="Deploy on Platform.sh" width="180px" />

				</a>

				We also offer a wide range of dedicated guides to all Meilisearch features, such as [filtering](https://www.meilisearch.com/docs/learn/fine_tuning_results/filtering?utm_campaign=oss&utm_source=github&utm_medium=meilisearch&utm_content=advanced), [sorting](https://www.meilisearch.com/docs/learn/fine_tuning_results/sorting?utm_campaign=oss&utm_source=github&utm_medium=meilisearch&utm_content=advanced), [geosearch](https://www.meilisearch.com/docs/learn/fine_tuning_results/geosearch?utm_campaign=oss&utm_source=github&utm_medium=meilisearch&utm_content=advanced), [API keys](https://www.meilisearch.com/docs/learn/security/master_api_keys?utm_campaign=oss&utm_source=github&utm_medium=meilisearch&utm_content=advanced), and [tenant tokens](https://www.meilisearch.com/docs/learn/security/tenant_tokens?utm_campaign=oss&utm_source=github&utm_medium=meilisearch&utm_content=advanced).

				#### APT (Debian & Ubuntu)

				Finally, for more in-depth information, refer to our articles explaining fundamental Meilisearch concepts such as [documents](https://www.meilisearch.com/docs/learn/core_concepts/documents?utm_campaign=oss&utm_source=github&utm_medium=meilisearch&utm_content=advanced) and [indexes](https://www.meilisearch.com/docs/learn/core_concepts/indexes?utm_campaign=oss&utm_source=github&utm_medium=meilisearch&utm_content=advanced).

				```bash

				echo "deb [trusted=yes] https://apt.fury.io/meilisearch/ /" > /etc/apt/sources.list.d/fury.list

				apt update && apt install meilisearch-http

				meilisearch

				```

				## 📊 Telemetry

				#### Download the binary (Linux & Mac OS)

				Meilisearch collects **anonymized** data from users to help us improve our product. You can [deactivate this](https://www.meilisearch.com/docs/learn/what_is_meilisearch/telemetry?utm_campaign=oss&utm_source=github&utm_medium=meilisearch&utm_content=telemetry#how-to-disable-data-collection) whenever you want.

				```bash

				curl -L https://install.meilisearch.com | sh

				./meilisearch

				```

				To request deletion of collected data, please write to us at [privacy@meilisearch.com](mailto:privacy@meilisearch.com). Don't forget to include your `Instance UID` in the message, as this helps us quickly find and delete your data.

				#### Compile and run it from sources

				If you want to know more about the kind of data we collect and what we use it for, check the [telemetry section](https://www.meilisearch.com/docs/learn/what_is_meilisearch/telemetry?utm_campaign=oss&utm_source=github&utm_medium=meilisearch&utm_content=telemetry#how-to-disable-data-collection) of our documentation.

				If you have the latest stable Rust toolchain installed on your local system, clone the repository and change it to your working directory.

				## 📫 Get in touch!

				```bash

				git clone https://github.com/meilisearch/meilisearch.git

				cd meilisearch

				cargo run --release

				```

				Meilisearch is a search engine created by [Meili](https://www.welcometothejungle.com/en/companies/meilisearch), a software development company based in France and with team members all over the world. Want to know more about us? [Check out our blog!](https://blog.meilisearch.com/?utm_campaign=oss&utm_source=github&utm_medium=meilisearch&utm_content=contact)

				### Create an Index and Upload Some Documents

				🗞 [Subscribe to our newsletter](https://meilisearch.us2.list-manage.com/subscribe?u=27870f7b71c908a8b359599fb&id=79582d828e) if you don't want to miss any updates! We promise we won't clutter your mailbox: we only send one edition every two months.

				Let's create an index! If you need a sample dataset, use [this movie database](https://www.notion.so/meilisearch/A-movies-dataset-to-test-Meili-1cbf7c9cfa4247249c40edfa22d7ca87#b5ae399b81834705ba5420ac70358a65). You can also find it in the `datasets/` directory.

				💌 Want to make a suggestion or give feedback? Here are some of the channels where you can reach us:

				```bash

				curl -L 'https://bit.ly/2PAcw9l' -o movies.json

				```

				- For feature requests, please visit our [product repository](https://github.com/meilisearch/product/discussions)

				- Found a bug? Open an [issue](https://github.com/meilisearch/meilisearch/issues)!

				- Want to be part of our Discord community? [Join us!](https://discord.meilisearch.com/?utm_campaign=oss&utm_source=github&utm_medium=meilisearch&utm_content=contact)

				Now, you're ready to index some data.

				Thank you for your support!

				```bash

				curl -i -X POST 'http://127.0.0.1:7700/indexes/movies/documents' \

				  --header 'content-type: application/json' \

				  --data-binary @movies.json

				```

				## 👩‍💻 Contributing

				### Search for Documents

				Meilisearch is, and will always be, open-source! If you want to contribute to the project, please take a look at [our contribution guidelines](CONTRIBUTING.md).

				#### In command line

				## 📦 Versioning

				The search engine is now aware of your documents and can serve those via an HTTP server.

				Meilisearch releases and their associated binaries are available [in this GitHub page](https://github.com/meilisearch/meilisearch/releases).

				The [`jq` command-line tool](https://stedolan.github.io/jq/) can greatly help you read the server responses.

				The binaries are versioned following [SemVer conventions](https://semver.org/). To know more, read our [versioning policy](https://github.com/meilisearch/engine-team/blob/main/resources/versioning-policy.md).

				```bash

				curl 'http://127.0.0.1:7700/indexes/movies/search?q=botman+robin&limit=2' | jq

				```

				```json

				{

				  "hits": [

				    {

				      "id": "415",

				      "title": "Batman & Robin",

				      "poster": "https://image.tmdb.org/t/p/w1280/79AYCcxw3kSKbhGpx1LiqaCAbwo.jpg",

				      "overview": "Along with crime-fighting partner Robin and new recruit Batgirl, Batman battles the dual threat of frosty genius Mr. Freeze and homicidal horticulturalist Poison Ivy. Freeze plans to put Gotham City on ice, while Ivy tries to drive a wedge between the dynamic duo.",

				      "release_date": 866768400

				    },

				    {

				      "id": "411736",

				      "title": "Batman: Return of the Caped Crusaders",

				      "poster": "https://image.tmdb.org/t/p/w1280/GW3IyMW5Xgl0cgCN8wu96IlNpD.jpg",

				      "overview": "Adam West and Burt Ward returns to their iconic roles of Batman and Robin. Featuring the voices of Adam West, Burt Ward, and Julie Newmar, the film sees the superheroes going up against classic villains like The Joker, The Riddler, The Penguin and Catwoman, both in Gotham City… and in space.",

				      "release_date": 1475888400

				    }

				  ],

				  "nbHits": 8,

				  "exhaustiveNbHits": false,

				  "query": "botman robin",

				  "limit": 2,

				  "offset": 0,

				  "processingTimeMs": 2

				}

				```

				#### Use the Web Interface

				We also deliver an **out-of-the-box [web interface](https://github.com/meilisearch/mini-dashboard)** in which you can test Meilisearch interactively.

				You can access the web interface in your web browser at the root of the server. The default URL is [http://127.0.0.1:7700](http://127.0.0.1:7700). All you need to do is open your web browser and enter Meilisearch’s address to visit it. This will lead you to a web page with a search bar that will allow you to search in the selected index.

				| [See the gif above](#demo)

				## Documentation

				Now that your Meilisearch server is up and running, you can learn more about how to tune your search engine in [the documentation](https://docs.meilisearch.com).

				## Contributing

				Hey! We're glad you're thinking about contributing to Meilisearch! Feel free to pick an [issue labeled as `good first issue`](https://github.com/meilisearch/meilisearch/issues?q=is%3Aissue+is%3Aopen+label%3A%22good+first+issue%22), and to ask any question you need. Some points might not be clear and we are available to help you!

				Also, we recommend following the [CONTRIBUTING](./CONTRIBUTING.md) to create your PR.

				## Core engine and tokenizer

				The code in this repository is only concerned with managing multiple indexes, handling the update store, and exposing an HTTP API.

				Search and indexation are the domain of our core engine, [`milli`](https://github.com/meilisearch/milli), while tokenization is handled by [our `tokenizer` library](https://github.com/meilisearch/tokenizer/).

				## Telemetry

				Meilisearch collects anonymous data regarding general usage.

				This helps us better understand developers' usage of Meilisearch features.

				To find out more on what information we're retrieving, please see our documentation on [Telemetry](https://docs.meilisearch.com/learn/what_is_meilisearch/telemetry.html).

				This program is optional, you can disable these analytics by using the `MEILI_NO_ANALYTICS` env variable.

				## Feature request

				The feature requests are not managed in this repository. Please visit our [dedicated repository](https://github.com/meilisearch/product) to see our work about the Meilisearch product.

				If you have a feature request or any feedback about an existing feature, please open [a discussion](https://github.com/meilisearch/product/discussions).

				Also, feel free to participate in the current discussions, we are looking forward to reading your comments.

				## 💌 Contact

				Please visit [this page](https://docs.meilisearch.com/learn/what_is_meilisearch/contact.html#contact-us).

				Meilisearch is developed by [Meili](https://www.meilisearch.com), a young company. To know more about us, you can [read our blog](https://blog.meilisearch.com). Any suggestion or feedback is highly appreciated. Thank you for your support!

				Differently from the binaries, crates in this repository are not currently available on [crates.io](https://crates.io/) and do not follow [SemVer conventions](https://semver.org).

									
										2

SECURITY.md
									
												View File
												
				@@ -4,7 +4,7 @@ Meilisearch takes the security of our software products and services seriously.

				If you believe you have found a security vulnerability in any Meilisearch-owned repository, please report it to us as described below.

				## Suported versions

				## Supported versions

				As long as we are pre-v1.0, only the latest version of Meilisearch will be supported with security updates.

BIN
assets/demo-dark.gif Normal file

View File

Binary file not shown.

After

Width: | Height: | Size: 2.8 MiB

BIN
assets/demo-light.gif Normal file

View File

Binary file not shown.

After

Width: | Height: | Size: 1.7 MiB

1393

assets/grafana-dashboard.json Normal file

View File

File diff suppressed because it is too large Load Diff

BIN
assets/integrations.png Normal file

View File

Binary file not shown.

After

Width: | Height: | Size: 799 KiB

30

assets/meilisearch-logo-dark.svg Normal file

View File

@@ -0,0 +1,30 @@
 <svg width="495" height="74" viewBox="0 0 495 74" fill="none" xmlns="http://www.w3.org/2000/svg">
 <path d="M181.842 42.5349C181.842 37.6137 184.201 34.715 188.718 34.715C192.965 34.715 194.381 37.7486 194.381 41.6585V62.6238H203.953V40.5799C203.953 32.3556 199.639 26.4907 191.145 26.4907C186.089 26.4907 182.516 28.0412 179.415 31.4792C177.393 28.3782 173.955 26.4907 169.168 26.4907C164.112 26.4907 160.607 28.5805 158.989 31.614V27.2996H150.158V62.6238H159.731V42.3326C159.731 37.6137 162.157 34.715 166.607 34.715C170.854 34.715 172.269 37.7486 172.269 41.6585V62.6238H181.842V42.5349Z" fill="white"/>
 <path d="M243.245 47.7256C243.245 47.7256 243.379 46.4448 243.379 44.8943C243.379 34.4454 236.301 26.4907 225.852 26.4907C215.403 26.4907 208.123 34.4454 208.123 44.8943C208.123 55.7477 215.471 63.4327 225.92 63.4327C234.077 63.4327 240.548 58.5116 242.638 51.3659H232.998C231.852 53.9276 229.088 55.2084 226.189 55.2084C221.403 55.2084 218.302 52.5793 217.628 47.7256H243.245ZM225.785 34.1757C230.234 34.1757 233.133 36.8722 233.807 40.8495H217.763C218.572 36.8048 221.403 34.1757 225.785 34.1757Z" fill="white"/>
 <path d="M244.791 35.524H249.038V62.6238H258.61V27.2996H244.791V35.524ZM253.824 22.7156C257.195 22.7156 259.622 20.3561 259.622 16.9855C259.622 13.6149 257.195 11.188 253.824 11.188C250.454 11.188 248.027 13.6149 248.027 16.9855C248.027 20.3561 250.454 22.7156 253.824 22.7156Z" fill="white"/>
 <path d="M278.432 54.3995C278.163 54.3995 277.758 54.4669 277.152 54.4669C274.994 54.4669 274.725 53.4557 274.725 51.9726V12.0644H265.152V52.6467C265.152 59.6576 267.849 62.7586 275.466 62.7586C276.747 62.7586 277.96 62.6238 278.432 62.5564V54.3995Z" fill="white"/>
 <path d="M279.521 35.524H283.768V62.6238H293.341V27.2996H279.521V35.524ZM288.555 22.7156C291.925 22.7156 294.352 20.3561 294.352 16.9855C294.352 13.6149 291.925 11.188 288.555 11.188C285.184 11.188 282.757 13.6149 282.757 16.9855C282.757 20.3561 285.184 22.7156 288.555 22.7156Z" fill="white"/>
 <path d="M312.557 62.9937C321.86 62.9937 326.242 58.0726 326.242 52.8819C326.242 38.4556 305.007 46.4777 305.007 36.9725C305.007 33.8716 307.636 31.2425 312.962 31.2425C318.422 31.2425 320.984 34.2086 321.388 37.9163H326.175C325.77 33.2648 322.602 27.0629 313.097 27.0629C304.94 27.0629 300.356 31.9166 300.356 37.1748C300.356 51.264 321.591 43.1745 321.591 53.0167C321.591 56.4547 318.355 58.8142 312.557 58.8142C306.625 58.8142 303.659 55.848 303.322 51.4662H298.468C298.873 57.4659 302.648 62.9937 312.557 62.9937Z" fill="white"/>
 <path d="M364.256 46.4103C364.256 46.4103 364.324 45.3317 364.324 44.5901C364.324 34.8827 358.054 27.0629 347.808 27.0629C337.494 27.0629 330.955 35.4894 330.955 44.9946C330.955 54.6346 337.022 62.9937 347.875 62.9937C356.032 62.9937 361.695 58.0052 363.717 51.4662H358.729C357.245 55.6458 353.201 58.6794 347.943 58.6794C340.729 58.6794 336.213 53.3538 335.741 46.4103H364.256ZM347.808 31.3773C354.549 31.3773 358.931 35.8939 359.538 42.5004H335.876C336.685 36.1636 341.134 31.3773 347.808 31.3773Z" fill="white"/>
 <path d="M394.037 45.871V49.1068C394.037 54.9717 389.79 59.0164 381.634 59.0164C376.578 59.0164 373.814 56.9266 373.814 52.41C373.814 50.118 374.892 48.3652 376.578 47.4215C378.33 46.4777 380.69 45.871 394.037 45.871ZM381.094 62.9937C387.027 62.9937 391.813 61.1062 394.24 57.1963V62.1848H398.824V39.7364C398.824 32.1188 394.442 27.0629 384.532 27.0629C375.027 27.0629 370.848 31.8492 369.971 37.9837H374.623C375.566 33.13 379.274 31.1751 384.33 31.1751C390.802 31.1751 394.037 33.8716 394.037 39.669V41.8936C383.184 41.8936 378.667 42.0959 375.297 43.4441C371.387 44.9946 369.095 48.4327 369.095 52.5448C369.095 58.5445 372.937 62.9937 381.094 62.9937Z" fill="white"/>
 <path d="M424.991 27.6022C424.991 27.6022 424.182 27.5348 423.845 27.5348C417.509 27.5348 414.138 30.838 412.857 33.1974V27.8718H408.273V62.1848H413.059V42.7026C413.059 35.5569 417.441 32.0514 423.306 32.0514C424.182 32.0514 424.991 32.1188 424.991 32.1188V27.6022Z" fill="white"/>
 <path d="M425.809 45.062C425.809 54.4324 432.28 62.9937 442.729 62.9937C452.032 62.9937 457.425 56.7918 458.773 49.9831H453.92C452.504 55.3087 448.594 58.6794 442.729 58.6794C435.516 58.6794 430.662 52.9493 430.662 45.062C430.662 37.1073 435.516 31.3773 442.729 31.3773C448.594 31.3773 452.504 34.7479 453.92 40.0735H458.773C457.425 33.2648 452.032 27.0629 442.729 27.0629C432.28 27.0629 425.809 35.6243 425.809 45.062Z" fill="white"/>
 <path d="M470.041 11.6254H465.255V62.1848H470.041V41.8936C470.041 34.8827 474.558 31.2425 480.355 31.2425C486.49 31.2425 489.389 35.0176 489.389 41.2195V62.1848H494.175V40.2757C494.175 32.6581 489.658 27.0629 481.164 27.0629C474.76 27.0629 471.255 30.5683 470.041 32.6581V11.6254Z" fill="white"/>
 <path d="M0.825012 73.993L24.0688 14.5224C27.3443 6.14179 35.4223 0.625977 44.4203 0.625977H58.4336L35.1899 60.0966C31.9143 68.4772 23.8363 73.993 14.8384 73.993H0.825012Z" fill="url(#paint0_linear_0_3)"/>
 <path d="M34.9246 73.9932L58.1684 14.5226C61.444 6.14197 69.5219 0.626152 78.5199 0.626152H92.5333L69.2895 60.0968C66.014 68.4774 57.936 73.9932 48.938 73.9932H34.9246Z" fill="url(#paint1_linear_0_3)"/>
 <path d="M69.0262 73.9932L92.27 14.5226C95.5456 6.14197 103.624 0.626152 112.622 0.626152H126.635L103.391 60.0968C100.116 68.4774 92.0376 73.9932 83.0396 73.9932H69.0262Z" fill="url(#paint2_linear_0_3)"/>
 <defs>
 <linearGradient id="paint0_linear_0_3" x1="126.635" y1="-4.97799" x2="0.825008" y2="66.0978" gradientUnits="userSpaceOnUse">
 <stop stop-color="#FF5CAA"/>
 <stop offset="1" stop-color="#FF4E62"/>
 </linearGradient>
 <linearGradient id="paint1_linear_0_3" x1="126.635" y1="-4.97799" x2="0.825008" y2="66.0978" gradientUnits="userSpaceOnUse">
 <stop stop-color="#FF5CAA"/>
 <stop offset="1" stop-color="#FF4E62"/>
 </linearGradient>
 <linearGradient id="paint2_linear_0_3" x1="126.635" y1="-4.97799" x2="0.825008" y2="66.0978" gradientUnits="userSpaceOnUse">
 <stop stop-color="#FF5CAA"/>
 <stop offset="1" stop-color="#FF4E62"/>
 </linearGradient>
 </defs>
 </svg>

After

Width: | Height: | Size: 5.8 KiB

30

assets/meilisearch-logo-light.svg Normal file

View File

@@ -0,0 +1,30 @@
 <svg width="495" height="74" viewBox="0 0 495 74" fill="none" xmlns="http://www.w3.org/2000/svg">
 <path d="M181.84 42.5347C181.84 37.6136 184.199 34.7149 188.716 34.7149C192.963 34.7149 194.378 37.7484 194.378 41.6584V62.6237H203.951V40.5798C203.951 32.3554 199.637 26.4906 191.143 26.4906C186.087 26.4906 182.514 28.041 179.413 31.4791C177.39 28.3781 173.952 26.4906 169.166 26.4906C164.11 26.4906 160.605 28.5804 158.987 31.6139V27.2995H150.156V62.6237H159.728V42.3325C159.728 37.6136 162.155 34.7149 166.604 34.7149C170.851 34.7149 172.267 37.7484 172.267 41.6584V62.6237H181.84V42.5347Z" fill="#21004B"/>
 <path d="M243.242 47.7255C243.242 47.7255 243.377 46.4447 243.377 44.8942C243.377 34.4452 236.299 26.4906 225.85 26.4906C215.401 26.4906 208.12 34.4452 208.12 44.8942C208.12 55.7476 215.468 63.4326 225.917 63.4326C234.074 63.4326 240.546 58.5115 242.636 51.3658H232.996C231.85 53.9274 229.086 55.2083 226.187 55.2083C221.401 55.2083 218.3 52.5792 217.626 47.7255H243.242ZM225.783 34.1756C230.232 34.1756 233.131 36.8721 233.805 40.8494H217.76C218.569 36.8047 221.401 34.1756 225.783 34.1756Z" fill="#21004B"/>
 <path d="M244.789 35.5238H249.036V62.6237H258.608V27.2995H244.789V35.5238ZM253.822 22.7155C257.193 22.7155 259.619 20.356 259.619 16.9854C259.619 13.6148 257.193 11.1879 253.822 11.1879C250.451 11.1879 248.024 13.6148 248.024 16.9854C248.024 20.356 250.451 22.7155 253.822 22.7155Z" fill="#21004B"/>
 <path d="M278.43 54.3993C278.16 54.3993 277.756 54.4667 277.149 54.4667C274.992 54.4667 274.722 53.4556 274.722 51.9725V12.0643H265.15V52.6466C265.15 59.6575 267.846 62.7585 275.464 62.7585C276.745 62.7585 277.958 62.6237 278.43 62.5562V54.3993Z" fill="#21004B"/>
 <path d="M279.519 35.5238H283.766V62.6237H293.339V27.2995H279.519V35.5238ZM288.553 22.7155C291.923 22.7155 294.35 20.356 294.35 16.9854C294.35 13.6148 291.923 11.1879 288.553 11.1879C285.182 11.1879 282.755 13.6148 282.755 16.9854C282.755 20.356 285.182 22.7155 288.553 22.7155Z" fill="#21004B"/>
 <path d="M312.557 62.9939C321.86 62.9939 326.242 58.0728 326.242 52.882C326.242 38.4557 305.007 46.4778 305.007 36.9726C305.007 33.8717 307.636 31.2426 312.962 31.2426C318.422 31.2426 320.984 34.2087 321.388 37.9164H326.175C325.77 33.265 322.602 27.063 313.097 27.063C304.94 27.063 300.356 31.9167 300.356 37.1749C300.356 51.2641 321.591 43.1746 321.591 53.0168C321.591 56.4548 318.355 58.8143 312.557 58.8143C306.625 58.8143 303.659 55.8481 303.322 51.4663H298.468C298.872 57.466 302.648 62.9939 312.557 62.9939Z" fill="#21004B"/>
 <path d="M364.256 46.4104C364.256 46.4104 364.324 45.3318 364.324 44.5903C364.324 34.8829 358.054 27.063 347.808 27.063C337.494 27.063 330.955 35.4896 330.955 44.9947C330.955 54.6347 337.022 62.9939 347.875 62.9939C356.032 62.9939 361.695 58.0053 363.717 51.4663H358.728C357.245 55.6459 353.201 58.6795 347.942 58.6795C340.729 58.6795 336.213 53.3539 335.741 46.4104H364.256ZM347.808 31.3774C354.549 31.3774 358.931 35.894 359.537 42.5005H335.876C336.685 36.1637 341.134 31.3774 347.808 31.3774Z" fill="#21004B"/>
 <path d="M394.037 45.8711V49.1069C394.037 54.9718 389.79 59.0165 381.633 59.0165C376.578 59.0165 373.814 56.9267 373.814 52.4101C373.814 50.1181 374.892 48.3654 376.578 47.4216C378.33 46.4778 380.69 45.8711 394.037 45.8711ZM381.094 62.9939C387.026 62.9939 391.813 61.1063 394.24 57.1964V62.1849H398.824V39.7366C398.824 32.1189 394.442 27.063 384.532 27.063C375.027 27.063 370.847 31.8493 369.971 37.9838H374.623C375.566 33.1301 379.274 31.1752 384.33 31.1752C390.802 31.1752 394.037 33.8717 394.037 39.6691V41.8938C383.184 41.8938 378.667 42.096 375.297 43.4442C371.387 44.9947 369.095 48.4328 369.095 52.5449C369.095 58.5446 372.937 62.9939 381.094 62.9939Z" fill="#21004B"/>
 <path d="M424.991 27.6023C424.991 27.6023 424.182 27.5349 423.845 27.5349C417.508 27.5349 414.138 30.8381 412.857 33.1975V27.872H408.273V62.1849H413.059V42.7027C413.059 35.557 417.441 32.0515 423.306 32.0515C424.182 32.0515 424.991 32.1189 424.991 32.1189V27.6023Z" fill="#21004B"/>
 <path d="M425.809 45.0621C425.809 54.4325 432.28 62.9939 442.729 62.9939C452.032 62.9939 457.425 56.7919 458.773 49.9832H453.92C452.504 55.3088 448.594 58.6795 442.729 58.6795C435.516 58.6795 430.662 52.9494 430.662 45.0621C430.662 37.1075 435.516 31.3774 442.729 31.3774C448.594 31.3774 452.504 34.748 453.92 40.0736H458.773C457.425 33.265 452.032 27.063 442.729 27.063C432.28 27.063 425.809 35.6244 425.809 45.0621Z" fill="#21004B"/>
 <path d="M470.041 11.6255H465.255V62.1849H470.041V41.8938C470.041 34.8829 474.558 31.2426 480.355 31.2426C486.49 31.2426 489.389 35.0177 489.389 41.2196V62.1849H494.175V40.2759C494.175 32.6582 489.658 27.063 481.164 27.063C474.76 27.063 471.255 30.5685 470.041 32.6582V11.6255Z" fill="#21004B"/>
 <path d="M0.824951 73.993L24.0688 14.5224C27.3443 6.14179 35.4223 0.625977 44.4202 0.625977H58.4336L35.1898 60.0966C31.9143 68.4772 23.8363 73.993 14.8383 73.993H0.824951Z" fill="url(#paint0_linear_0_15)"/>
 <path d="M34.9246 73.9932L58.1684 14.5226C61.4439 6.14197 69.5219 0.626152 78.5199 0.626152H92.5332L69.2894 60.0968C66.0139 68.4774 57.9359 73.9932 48.9379 73.9932H34.9246Z" fill="url(#paint1_linear_0_15)"/>
 <path d="M69.0262 73.9932L92.27 14.5226C95.5455 6.14197 103.623 0.626152 112.621 0.626152H126.635L103.391 60.0968C100.115 68.4774 92.0375 73.9932 83.0395 73.9932H69.0262Z" fill="url(#paint2_linear_0_15)"/>
 <defs>
 <linearGradient id="paint0_linear_0_15" x1="126.635" y1="-4.97799" x2="0.824952" y2="66.0978" gradientUnits="userSpaceOnUse">
 <stop stop-color="#FF5CAA"/>
 <stop offset="1" stop-color="#FF4E62"/>
 </linearGradient>
 <linearGradient id="paint1_linear_0_15" x1="126.635" y1="-4.97799" x2="0.824952" y2="66.0978" gradientUnits="userSpaceOnUse">
 <stop stop-color="#FF5CAA"/>
 <stop offset="1" stop-color="#FF4E62"/>
 </linearGradient>
 <linearGradient id="paint2_linear_0_15" x1="126.635" y1="-4.97799" x2="0.824952" y2="66.0978" gradientUnits="userSpaceOnUse">
 <stop stop-color="#FF5CAA"/>
 <stop offset="1" stop-color="#FF4E62"/>
 </linearGradient>
 </defs>
 </svg>

After

Width: | Height: | Size: 5.9 KiB

6

assets/milli-logo.svg Normal file

View File

@@ -0,0 +1,6 @@
 <svg width="277" height="236" viewBox="0 0 277 236" fill="none" xmlns="http://www.w3.org/2000/svg">
 <path fill-rule="evenodd" clip-rule="evenodd" d="M213.085 190L242.907 86H276.196L246.375 190H213.085Z" fill="#494949"/>
 <path fill-rule="evenodd" clip-rule="evenodd" d="M0 190L29.8215 86H63.1111L33.2896 190H0Z" fill="#494949"/>
 <path fill-rule="evenodd" clip-rule="evenodd" d="M124.986 0L57.5772 235.083L60.7752 236H90.6038L158.276 0H124.986Z" fill="#494949"/>
 <path fill-rule="evenodd" clip-rule="evenodd" d="M195.273 0L127.601 236H160.891L228.563 0H195.273Z" fill="#494949"/>
 </svg>

After

Width: | Height: | Size: 585 B

BIN
assets/profiling-example.png Normal file

View File

Binary file not shown.

After

Width: | Height: | Size: 1.2 MiB

									
										19

assets/prometheus-basic-scraper.yml
									
										Normal file
									
												View File
												
				@@ -0,0 +1,19 @@

				global:

				  scrape_interval:     15s # By default, scrape targets every 15 seconds.

				  # Attach these labels to any time series or alerts when communicating with

				  # external systems (federation, remote storage, Alertmanager).

				  external_labels:

				    monitor: 'codelab-monitor'

				# A scrape configuration containing exactly one endpoint to scrape:

				# Here it's Prometheus itself.

				scrape_configs:

				  # The job name is added as a label `job=<job_name>` to any timeseries scraped from this config.

				  - job_name: 'meilisearch'

				    # Override the global default and scrape targets from this job every 5 seconds.

				    scrape_interval: 5s

				    static_configs:

				      - targets: ['localhost:7700']

1

benchmarks/.gitignore vendored Normal file

View File

				`@@ -0,0 +1 @@`
				`benches/datasets_paths.rs`

									
										50

benchmarks/Cargo.toml
									
										Normal file
									
												View File
												
				@@ -0,0 +1,50 @@

				[package]

				name = "benchmarks"

				publish = false

				version.workspace = true

				authors.workspace = true

				description.workspace = true

				homepage.workspace = true

				readme.workspace = true

				edition.workspace = true

				license.workspace = true

				[dependencies]

				anyhow = "1.0.79"

				csv = "1.3.0"

				milli = { path = "../milli" }

				mimalloc = { version = "0.1.39", default-features = false }

				serde_json = { version = "1.0.111", features = ["preserve_order"] }

				[dev-dependencies]

				criterion = { version = "0.5.1", features = ["html_reports"] }

				rand = "0.8.5"

				rand_chacha = "0.3.1"

				roaring = "0.10.2"

				[build-dependencies]

				anyhow = "1.0.79"

				bytes = "1.5.0"

				convert_case = "0.6.0"

				flate2 = "1.0.28"

				reqwest = { version = "0.11.23", features = ["blocking", "rustls-tls"], default-features = false }

				[features]

				default = ["milli/all-tokenizations"]

				[[bench]]

				name = "search_songs"

				harness = false

				[[bench]]

				name = "search_wiki"

				harness = false

				[[bench]]

				name = "search_geo"

				harness = false

				[[bench]]

				name = "indexing"

				harness = false

									
										138

benchmarks/README.md
									
										Normal file
									
												View File
												
				@@ -0,0 +1,138 @@

				Benchmarks

				==========

				## TOC

				- [Run the benchmarks](#run-the-benchmarks)

				- [Comparison between benchmarks](#comparison-between-benchmarks)

				- [Datasets](#datasets)

				## Run the benchmarks

				### On our private server

				The Meili team has self-hosted his own GitHub runner to run benchmarks on our dedicated bare metal server.

				To trigger the benchmark workflow:

				- Go to the `Actions` tab of this repository.

				- Select the `Benchmarks` workflow on the left.

				- Click on `Run workflow` in the blue banner.

				- Select the branch on which you want to run the benchmarks and select the dataset you want (default: `songs`).

				- Finally, click on `Run workflow`.

				This GitHub workflow will run the benchmarks and push the `critcmp` report to a DigitalOcean Space (= S3).

				The name of the uploaded file is displayed in the workflow.

				_[More about critcmp](https://github.com/BurntSushi/critcmp)._

				💡 To compare the just-uploaded benchmark with another one, check out the [next section](#comparison-between-benchmarks).

				### On your machine

				To run all the benchmarks (~5h):

				```bash

				cargo bench

				```

				To run only the `search_songs` (~1h), `search_wiki` (~3h), `search_geo` (~20m) or `indexing` (~2h) benchmark:

				```bash

				cargo bench --bench <dataset name>

				```

				By default, the benchmarks will be downloaded and uncompressed automatically in the target directory.<br>

				If you don't want to download the datasets every time you update something on the code, you can specify a custom directory with the environment variable `MILLI_BENCH_DATASETS_PATH`:

				```bash

				mkdir ~/datasets

				MILLI_BENCH_DATASETS_PATH=~/datasets cargo bench --bench search_songs # the four datasets are downloaded

				touch build.rs

				MILLI_BENCH_DATASETS_PATH=~/datasets cargo bench --bench songs # the code is compiled again but the datasets are not downloaded

				```

				## Comparison between benchmarks

				The benchmark reports we push are generated with `critcmp`. Thus, we use `critcmp` to show the result of a benchmark, or compare results between multiple benchmarks.

				We provide a script to download and display the comparison report.

				Requirements:

				- `grep`

				- `curl`

				- [`critcmp`](https://github.com/BurntSushi/critcmp)

				List the available file in the DO Space:

				```bash

				./benchmarks/script/list.sh

				```

				```bash

				songs_main_09a4321.json

				songs_geosearch_24ec456.json

				search_songs_main_cb45a10b.json

				```

				Run the comparison script:

				```bash

				# we get the result of ONE benchmark, this give you an idea of how much time an operation took

				./benchmarks/scripts/compare.sh son songs_geosearch_24ec456.json

				# we compare two benchmarks

				./benchmarks/scripts/compare.sh songs_main_09a4321.json songs_geosearch_24ec456.json

				# we compare three benchmarks

				./benchmarks/scripts/compare.sh songs_main_09a4321.json songs_geosearch_24ec456.json search_songs_main_cb45a10b.json

				```

				## Datasets

				The benchmarks uses the following datasets:

				- `smol-songs`

				- `smol-wiki`

				- `movies`

				- `smol-all-countries`

				### Songs

				`smol-songs` is a subset of the [`songs.csv` dataset](https://milli-benchmarks.fra1.digitaloceanspaces.com/datasets/songs.csv.gz).

				It was generated with this command:

				```bash

				xsv sample --seed 42 1000000 songs.csv -o smol-songs.csv

				```

				_[Download the generated `smol-songs` dataset](https://milli-benchmarks.fra1.digitaloceanspaces.com/datasets/smol-songs.csv.gz)._

				### Wiki

				`smol-wiki` is a subset of the [`wikipedia-articles.csv` dataset](https://milli-benchmarks.fra1.digitaloceanspaces.com/datasets/wiki-articles.csv.gz).

				It was generated with the following command:

				```bash

				xsv sample --seed 42 500000 wiki-articles.csv -o smol-wiki-articles.csv

				```

				_[Download the `smol-wiki` dataset](https://milli-benchmarks.fra1.digitaloceanspaces.com/datasets/smol-wiki-articles.csv.gz)._

				### Movies

				`movies` is a really small dataset we uses as our example in the [getting started](https://www.meilisearch.com/docs/learn/getting_started/quick_start)

				_[Download the `movies` dataset](https://www.meilisearch.com/movies.json)._

				### All Countries

				`smol-all-countries` is a subset of the [`all-countries.csv` dataset](https://milli-benchmarks.fra1.digitaloceanspaces.com/datasets/all-countries.csv.gz)

				It has been converted to jsonlines and then edited so it matches our format for the `_geo` field.

				It was generated with the following command:

				```bash

				bat all-countries.csv.gz | gunzip | xsv sample --seed 42 1000000 | csv2json-lite | sd '"latitude":"(.*?)","longitude":"(.*?)"' '"_geo": { "lat": $1, "lng": $2 }' | sd '\[|\]|,$' '' | gzip > smol-all-countries.jsonl.gz

				```

				_[Download the `smol-all-countries` dataset](https://milli-benchmarks.fra1.digitaloceanspaces.com/datasets/smol-all-countries.jsonl.gz)._

1347

benchmarks/benches/indexing.rs Normal file

View File

File diff suppressed because it is too large Load Diff

									
										122

benchmarks/benches/search_geo.rs
									
										Normal file
									
												View File
												
				@@ -0,0 +1,122 @@

				mod datasets_paths;

				mod utils;

				use criterion::{criterion_group, criterion_main};

				use milli::update::Settings;

				use utils::Conf;

				#[global_allocator]

				static ALLOC: mimalloc::MiMalloc = mimalloc::MiMalloc;

				fn base_conf(builder: &mut Settings) {

				    let displayed_fields =

				        ["geonameid", "name", "asciiname", "alternatenames", "_geo", "population"]

				            .iter()

				            .map(|s| s.to_string())

				            .collect();

				    builder.set_displayed_fields(displayed_fields);

				    let searchable_fields =

				        ["name", "alternatenames", "elevation"].iter().map(|s| s.to_string()).collect();

				    builder.set_searchable_fields(searchable_fields);

				    let filterable_fields =

				        ["_geo", "population", "elevation"].iter().map(|s| s.to_string()).collect();

				    builder.set_filterable_fields(filterable_fields);

				    let sortable_fields =

				        ["_geo", "population", "elevation"].iter().map(|s| s.to_string()).collect();

				    builder.set_sortable_fields(sortable_fields);

				}

				#[rustfmt::skip]

				const BASE_CONF: Conf = Conf {

				    dataset: datasets_paths::SMOL_ALL_COUNTRIES,

				    dataset_format: "jsonl",

				    queries: &[

				        "",

				    ],

				    configure: base_conf,

				    primary_key: Some("geonameid"),

				    ..Conf::BASE

				};

				fn bench_geo(c: &mut criterion::Criterion) {

				    #[rustfmt::skip]

				    let confs = &[

				        // A basic placeholder with no geo

				        utils::Conf {

				            group_name: "placeholder with no geo",

				            ..BASE_CONF

				        },

				        // Medium aglomeration: probably the most common usecase

				        utils::Conf {

				            group_name: "asc sort from Lille",

				            sort: Some(vec!["_geoPoint(50.62999333378238, 3.086269263384099):asc"]),

				            ..BASE_CONF

				        },

				        utils::Conf {

				            group_name: "desc sort from Lille",

				            sort: Some(vec!["_geoPoint(50.62999333378238, 3.086269263384099):desc"]),

				            ..BASE_CONF

				        },

				        // Big agglomeration: a lot of documents close to our point

				        utils::Conf {

				            group_name: "asc sort from Tokyo",

				            sort: Some(vec!["_geoPoint(35.749512532692144, 139.61664952543356):asc"]),

				            ..BASE_CONF

				        },

				        utils::Conf {

				            group_name: "desc sort from Tokyo",

				            sort: Some(vec!["_geoPoint(35.749512532692144, 139.61664952543356):desc"]),

				            ..BASE_CONF

				        },

				        // The furthest point from any civilization

				        utils::Conf {

				            group_name: "asc sort from Point Nemo",

				            sort: Some(vec!["_geoPoint(-48.87561645055408, -123.39275749319793):asc"]),

				            ..BASE_CONF

				        },

				        utils::Conf {

				            group_name: "desc sort from Point Nemo",

				            sort: Some(vec!["_geoPoint(-48.87561645055408, -123.39275749319793):desc"]),

				            ..BASE_CONF

				        },

				        // Filters

				        utils::Conf {

				            group_name: "filter of 100km from Lille",

				            filter: Some("_geoRadius(50.62999333378238, 3.086269263384099, 100000)"),

				            ..BASE_CONF

				        },

				        utils::Conf {

				            group_name: "filter of 1km from Lille",

				            filter: Some("_geoRadius(50.62999333378238, 3.086269263384099, 1000)"),

				            ..BASE_CONF

				        },

				        utils::Conf {

				            group_name: "filter of 100km from Tokyo",

				            filter: Some("_geoRadius(35.749512532692144, 139.61664952543356, 100000)"),

				            ..BASE_CONF

				        },

				        utils::Conf {

				            group_name: "filter of 1km from Tokyo",

				            filter: Some("_geoRadius(35.749512532692144, 139.61664952543356, 1000)"),

				            ..BASE_CONF

				        },

				        utils::Conf {

				            group_name: "filter of 100km from Point Nemo",

				            filter: Some("_geoRadius(-48.87561645055408, -123.39275749319793, 100000)"),

				            ..BASE_CONF

				        },

				        utils::Conf {

				            group_name: "filter of 1km from Point Nemo",

				            filter: Some("_geoRadius(-48.87561645055408, -123.39275749319793, 1000)"),

				            ..BASE_CONF

				        },

				    ];

				    utils::run_benches(c, confs);

				}

				criterion_group!(benches, bench_geo);

				criterion_main!(benches);

									
										196

benchmarks/benches/search_songs.rs
									
										Normal file
									
												View File
												
				@@ -0,0 +1,196 @@

				mod datasets_paths;

				mod utils;

				use criterion::{criterion_group, criterion_main};

				use milli::update::Settings;

				use utils::Conf;

				#[global_allocator]

				static ALLOC: mimalloc::MiMalloc = mimalloc::MiMalloc;

				fn base_conf(builder: &mut Settings) {

				    let displayed_fields =

				        ["id", "title", "album", "artist", "genre", "country", "released", "duration"]

				            .iter()

				            .map(|s| s.to_string())

				            .collect();

				    builder.set_displayed_fields(displayed_fields);

				    let searchable_fields = ["title", "album", "artist"].iter().map(|s| s.to_string()).collect();

				    builder.set_searchable_fields(searchable_fields);

				    let faceted_fields = ["released-timestamp", "duration-float", "genre", "country", "artist"]

				        .iter()

				        .map(|s| s.to_string())

				        .collect();

				    builder.set_filterable_fields(faceted_fields);

				}

				#[rustfmt::skip]

				const BASE_CONF: Conf = Conf {

				    dataset: datasets_paths::SMOL_SONGS,

				    queries: &[

				        "john ",             // 9097

				        "david ",            // 4794

				        "charles ",          // 1957

				        "david bowie ",      // 1200

				        "michael jackson ",  // 600

				        "thelonious monk ",  // 303

				        "charles mingus ",   // 142

				        "marcus miller ",    // 60

				        "tamo ",             // 13

				        "Notstandskomitee ", // 4

				    ],

				    configure: base_conf,

				    primary_key: Some("id"),

				    ..Conf::BASE

				};

				fn bench_songs(c: &mut criterion::Criterion) {

				    let default_criterion: Vec<String> =

				        milli::default_criteria().iter().map(|criteria| criteria.to_string()).collect();

				    let default_criterion = default_criterion.iter().map(|s| s.as_str());

				    let asc_default: Vec<&str> =

				        std::iter::once("released-timestamp:asc").chain(default_criterion.clone()).collect();

				    let desc_default: Vec<&str> =

				        std::iter::once("released-timestamp:desc").chain(default_criterion.clone()).collect();

				    let basic_with_quote: Vec<String> = BASE_CONF

				        .queries

				        .iter()

				        .map(|s| {

				            s.trim().split(' ').map(|s| format!(r#""{}""#, s)).collect::<Vec<String>>().join(" ")

				        })

				        .collect();

				    let basic_with_quote: &[&str] =

				        &basic_with_quote.iter().map(|s| s.as_str()).collect::<Vec<&str>>();

				    #[rustfmt::skip]

				    let confs = &[

				        /* first we bench each criterion alone */

				        utils::Conf {

				            group_name: "proximity",

				            queries: &[

				                "black saint sinner lady ",

				                "les dangeureuses 1960 ",

				                "The Disneyland Sing-Along Chorus ",

				                "Under Great Northern Lights ",

				                "7000 Danses Un Jour Dans Notre Vie ",

				            ],

				            criterion: Some(&["proximity"]),

				            optional_words: false,

				            ..BASE_CONF

				        },

				        utils::Conf {

				            group_name: "typo",

				            queries: &[

				                "mongus ",

				                "thelonius monk ",

				                "Disnaylande ",

				                "the white striper ",

				                "indochie ",

				                "indochien ",

				                "klub des loopers ",

				                "fear of the duck ",

				                "michel depech ",

				                "stromal ",

				                "dire straights ",

				                "Arethla Franklin ",

				            ],

				            criterion: Some(&["typo"]),

				            optional_words: false,

				            ..BASE_CONF

				        },

				        utils::Conf {

				            group_name: "words",

				            queries: &[

				                "the black saint and the sinner lady and the good doggo ", // four words to pop

				                "les liaisons dangeureuses 1793 ",                         // one word to pop

				                "The Disneyland Children's Sing-Alone song ",              // two words to pop

				                "seven nation mummy ",                                     // one word to pop

				                "7000 Danses / Le Baiser / je me trompe de mots ",         // four words to pop

				                "Bring Your Daughter To The Slaughter but now this is not part of the title ", // nine words to pop

				                "whathavenotnsuchforth and a good amount of words to pop to match the first one ", // 13

				            ],

				            criterion: Some(&["words"]),

				            ..BASE_CONF

				        },

				        utils::Conf {

				            group_name: "asc",

				            criterion: Some(&["released-timestamp:desc"]),

				            ..BASE_CONF

				        },

				        utils::Conf {

				            group_name: "desc",

				            criterion: Some(&["released-timestamp:desc"]),

				            ..BASE_CONF

				        },

				        /* then we bench the asc and desc criterion on top of the default criterion */

				        utils::Conf {

				            group_name: "asc + default",

				            criterion: Some(&asc_default[..]),

				            ..BASE_CONF

				        },

				        utils::Conf {

				            group_name: "desc + default",

				            criterion: Some(&desc_default[..]),

				            ..BASE_CONF

				        },

				        /* we bench the filters with the default request */

				        utils::Conf {

				            group_name: "basic filter: <=",

				            filter: Some("released-timestamp <= 946728000"), // year 2000

				            ..BASE_CONF

				        },

				        utils::Conf {

				            group_name: "basic filter: TO",

				            filter: Some("released-timestamp 946728000 TO 1262347200"), // year 2000 to 2010

				            ..BASE_CONF

				        },

				        utils::Conf {

				            group_name: "big filter",

				            filter: Some("released-timestamp != 1262347200 AND (NOT (released-timestamp = 946728000)) AND (duration-float = 1 OR (duration-float 1.1 TO 1.5 AND released-timestamp > 315576000))"),

				            ..BASE_CONF

				        },

				        /* the we bench some global / normal search with all the default criterion in the default

				         * order */

				        utils::Conf {

				            group_name: "basic placeholder",

				            queries: &[""],

				            ..BASE_CONF

				        },

				        utils::Conf {

				            group_name: "basic without quote",

				            queries: &BASE_CONF

				                .queries

				                .iter()

				                .map(|s| s.trim()) // we remove the space at the end of each request

				                .collect::<Vec<&str>>(),

				            ..BASE_CONF

				        },

				        utils::Conf {

				            group_name: "basic with quote",

				            queries: basic_with_quote,

				            ..BASE_CONF

				        },

				        utils::Conf {

				            group_name: "prefix search",

				            queries: &[

				                "s", // 500k+ results

				                "a", //

				                "b", //

				                "i", //

				                "x", // only 7k results

				            ],

				            ..BASE_CONF

				        },

				    ];

				    utils::run_benches(c, confs);

				}

				criterion_group!(benches, bench_songs);

				criterion_main!(benches);

									
										129

benchmarks/benches/search_wiki.rs
									
										Normal file
									
												View File
												
				@@ -0,0 +1,129 @@

				mod datasets_paths;

				mod utils;

				use criterion::{criterion_group, criterion_main};

				use milli::update::Settings;

				use utils::Conf;

				#[global_allocator]

				static ALLOC: mimalloc::MiMalloc = mimalloc::MiMalloc;

				fn base_conf(builder: &mut Settings) {

				    let displayed_fields = ["title", "body", "url"].iter().map(|s| s.to_string()).collect();

				    builder.set_displayed_fields(displayed_fields);

				    let searchable_fields = ["title", "body"].iter().map(|s| s.to_string()).collect();

				    builder.set_searchable_fields(searchable_fields);

				}

				#[rustfmt::skip]

				const BASE_CONF: Conf = Conf {

				    dataset: datasets_paths::SMOL_WIKI_ARTICLES,

				    queries: &[

				        "mingus ",        // 46 candidates

				        "miles davis ",   // 159

				        "rock and roll ", // 1007

				        "machine ",       // 3448

				        "spain ",         // 7002

				        "japan ",         // 10.593

				        "france ",        // 17.616

				        "film ",          // 24.959

				    ],

				    configure: base_conf,

				    ..Conf::BASE

				};

				fn bench_songs(c: &mut criterion::Criterion) {

				    let basic_with_quote: Vec<String> = BASE_CONF

				        .queries

				        .iter()

				        .map(|s| {

				            s.trim().split(' ').map(|s| format!(r#""{}""#, s)).collect::<Vec<String>>().join(" ")

				        })

				        .collect();

				    let basic_with_quote: &[&str] =

				        &basic_with_quote.iter().map(|s| s.as_str()).collect::<Vec<&str>>();

				    #[rustfmt::skip]

				    let confs = &[

				        /* first we bench each criterion alone */

				        utils::Conf {

				            group_name: "proximity",

				            queries: &[

				                "herald sings ",

				                "april paris ",

				                "tea two ",

				                "diesel engine ",

				            ],

				            criterion: Some(&["proximity"]),

				            optional_words: false,

				            ..BASE_CONF

				        },

				        utils::Conf {

				            group_name: "typo",

				            queries: &[

				                "migrosoft ",

				                "linax ",

				                "Disnaylande ",

				                "phytogropher ",

				                "nympalidea ",

				                "aritmetric ",

				                "the fronce ",

				                "sisan ",

				            ],

				            criterion: Some(&["typo"]),

				            optional_words: false,

				            ..BASE_CONF

				        },

				        utils::Conf {

				            group_name: "words",

				            queries: &[

				                "the black saint and the sinner lady and the good doggo ", // four words to pop, 27 results

				                "Kameya Tokujirō mingus monk ",                            // two words to pop, 55

				                "Ulrich Hensel meilisearch milli ",                        // two words to pop, 306

				                "Idaho Bellevue pizza ",                                   // one word to pop, 800

				                "Abraham machin ",                                         // one word to pop, 1141

				            ],

				            criterion: Some(&["words"]),

				            ..BASE_CONF

				        },

				        /* the we bench some global / normal search with all the default criterion in the default

				         * order */

				        utils::Conf {

				            group_name: "basic placeholder",

				            queries: &[""],

				            ..BASE_CONF

				        },

				        utils::Conf {

				            group_name: "basic without quote",

				            queries: &BASE_CONF

				                .queries

				                .iter()

				                .map(|s| s.trim()) // we remove the space at the end of each request

				                .collect::<Vec<&str>>(),

				            ..BASE_CONF

				        },

				        utils::Conf {

				            group_name: "basic with quote",

				            queries: basic_with_quote,

				            ..BASE_CONF

				        },

				        utils::Conf {

				            group_name: "prefix search",

				            queries: &[

				                "t", // 453k results

				                "c", // 405k

				                "g", // 318k

				                "j", // 227k

				                "q", // 71k

				                "x", // 17k

				            ],

				            ..BASE_CONF

				        },

				    ];

				    utils::run_benches(c, confs);

				}

				criterion_group!(benches, bench_songs);

				criterion_main!(benches);

									
										256

benchmarks/benches/utils.rs
									
										Normal file
									
												View File
												
				@@ -0,0 +1,256 @@

				#![allow(dead_code)]

				use std::fs::{create_dir_all, remove_dir_all, File};

				use std::io::{self, BufRead, BufReader, Cursor, Read, Seek};

				use std::num::ParseFloatError;

				use std::path::Path;

				use std::str::FromStr;

				use criterion::BenchmarkId;

				use milli::documents::{DocumentsBatchBuilder, DocumentsBatchReader};

				use milli::heed::EnvOpenOptions;

				use milli::update::{

				    IndexDocuments, IndexDocumentsConfig, IndexDocumentsMethod, IndexerConfig, Settings,

				};

				use milli::{Criterion, Filter, Index, Object, TermsMatchingStrategy};

				use serde_json::Value;

				pub struct Conf<'a> {

				    /// where we are going to create our database.mmdb directory

				    /// each benchmark will first try to delete it and then recreate it

				    pub database_name: &'a str,

				    /// the dataset to be used, it must be an uncompressed csv

				    pub dataset: &'a str,

				    /// The format of the dataset

				    pub dataset_format: &'a str,

				    pub group_name: &'a str,

				    pub queries: &'a [&'a str],

				    /// here you can change which criterion are used and in which order.

				    /// - if you specify something all the base configuration will be thrown out

				    /// - if you don't specify anything (None) the default configuration will be kept

				    pub criterion: Option<&'a [&'a str]>,

				    /// the last chance to configure your database as you want

				    pub configure: fn(&mut Settings),

				    pub filter: Option<&'a str>,

				    pub sort: Option<Vec<&'a str>>,

				    /// enable or disable the optional words on the query

				    pub optional_words: bool,

				    /// primary key, if there is None we'll auto-generate docids for every documents

				    pub primary_key: Option<&'a str>,

				}

				impl Conf<'_> {

				    pub const BASE: Self = Conf {

				        database_name: "benches.mmdb",

				        dataset_format: "csv",

				        dataset: "",

				        group_name: "",

				        queries: &[],

				        criterion: None,

				        configure: |_| (),

				        filter: None,

				        sort: None,

				        optional_words: true,

				        primary_key: None,

				    };

				}

				pub fn base_setup(conf: &Conf) -> Index {

				    match remove_dir_all(conf.database_name) {

				        Ok(_) => (),

				        Err(e) if e.kind() == std::io::ErrorKind::NotFound => (),

				        Err(e) => panic!("{}", e),

				    }

				    create_dir_all(conf.database_name).unwrap();

				    let mut options = EnvOpenOptions::new();

				    options.map_size(100 * 1024 * 1024 * 1024); // 100 GB

				    options.max_readers(10);

				    let index = Index::new(options, conf.database_name).unwrap();

				    let config = IndexerConfig::default();

				    let mut wtxn = index.write_txn().unwrap();

				    let mut builder = Settings::new(&mut wtxn, &index, &config);

				    if let Some(primary_key) = conf.primary_key {

				        builder.set_primary_key(primary_key.to_string());

				    }

				    if let Some(criterion) = conf.criterion {

				        builder.reset_filterable_fields();

				        builder.reset_criteria();

				        builder.reset_stop_words();

				        let criterion = criterion.iter().map(|s| Criterion::from_str(s).unwrap()).collect();

				        builder.set_criteria(criterion);

				    }

				    (conf.configure)(&mut builder);

				    builder.execute(|_| (), || false).unwrap();

				    wtxn.commit().unwrap();

				    let config = IndexerConfig::default();

				    let mut wtxn = index.write_txn().unwrap();

				    let indexing_config = IndexDocumentsConfig {

				        autogenerate_docids: conf.primary_key.is_none(),

				        update_method: IndexDocumentsMethod::ReplaceDocuments,

				        ..Default::default()

				    };

				    let builder =

				        IndexDocuments::new(&mut wtxn, &index, &config, indexing_config, |_| (), || false).unwrap();

				    let documents = documents_from(conf.dataset, conf.dataset_format);

				    let (builder, user_error) = builder.add_documents(documents).unwrap();

				    user_error.unwrap();

				    builder.execute().unwrap();

				    wtxn.commit().unwrap();

				    index

				}

				pub fn run_benches(c: &mut criterion::Criterion, confs: &[Conf]) {

				    for conf in confs {

				        let index = base_setup(conf);

				        let file_name = Path::new(conf.dataset).file_name().and_then(|f| f.to_str()).unwrap();

				        let name = format!("{}: {}", file_name, conf.group_name);

				        let mut group = c.benchmark_group(&name);

				        for &query in conf.queries {

				            group.bench_with_input(BenchmarkId::from_parameter(query), &query, |b, &query| {

				                b.iter(|| {

				                    let rtxn = index.read_txn().unwrap();

				                    let mut search = index.search(&rtxn);

				                    search.query(query).terms_matching_strategy(TermsMatchingStrategy::default());

				                    if let Some(filter) = conf.filter {

				                        let filter = Filter::from_str(filter).unwrap().unwrap();

				                        search.filter(filter);

				                    }

				                    if let Some(sort) = &conf.sort {

				                        let sort = sort.iter().map(|sort| sort.parse().unwrap()).collect();

				                        search.sort_criteria(sort);

				                    }

				                    let _ids = search.execute().unwrap();

				                });

				            });

				        }

				        group.finish();

				        index.prepare_for_closing().wait();

				    }

				}

				pub fn documents_from(filename: &str, filetype: &str) -> DocumentsBatchReader<impl BufRead + Seek> {

				    let reader = File::open(filename)

				        .unwrap_or_else(|_| panic!("could not find the dataset in: {}", filename));

				    let reader = BufReader::new(reader);

				    let documents = match filetype {

				        "csv" => documents_from_csv(reader).unwrap(),

				        "json" => documents_from_json(reader).unwrap(),

				        "jsonl" => documents_from_jsonl(reader).unwrap(),

				        otherwise => panic!("invalid update format {:?}", otherwise),

				    };

				    DocumentsBatchReader::from_reader(Cursor::new(documents)).unwrap()

				}

				fn documents_from_jsonl(reader: impl BufRead) -> anyhow::Result<Vec<u8>> {

				    let mut documents = DocumentsBatchBuilder::new(Vec::new());

				    for result in serde_json::Deserializer::from_reader(reader).into_iter::<Object>() {

				        let object = result?;

				        documents.append_json_object(&object)?;

				    }

				    documents.into_inner().map_err(Into::into)

				}

				fn documents_from_json(reader: impl BufRead) -> anyhow::Result<Vec<u8>> {

				    let mut documents = DocumentsBatchBuilder::new(Vec::new());

				    documents.append_json_array(reader)?;

				    documents.into_inner().map_err(Into::into)

				}

				fn documents_from_csv(reader: impl BufRead) -> anyhow::Result<Vec<u8>> {

				    let csv = csv::Reader::from_reader(reader);

				    let mut documents = DocumentsBatchBuilder::new(Vec::new());

				    documents.append_csv(csv)?;

				    documents.into_inner().map_err(Into::into)

				}

				enum AllowedType {

				    String,

				    Number,

				}

				fn parse_csv_header(header: &str) -> (String, AllowedType) {

				    // if there are several separators we only split on the last one.

				    match header.rsplit_once(':') {

				        Some((field_name, field_type)) => match field_type {

				            "string" => (field_name.to_string(), AllowedType::String),

				            "number" => (field_name.to_string(), AllowedType::Number),

				            // we may return an error in this case.

				            _otherwise => (header.to_string(), AllowedType::String),

				        },

				        None => (header.to_string(), AllowedType::String),

				    }

				}

				struct CSVDocumentDeserializer<R>

				where

				    R: Read,

				{

				    documents: csv::StringRecordsIntoIter<R>,

				    headers: Vec<(String, AllowedType)>,

				}

				impl<R: Read> CSVDocumentDeserializer<R> {

				    fn from_reader(reader: R) -> io::Result<Self> {

				        let mut records = csv::Reader::from_reader(reader);

				        let headers = records.headers()?.into_iter().map(parse_csv_header).collect();

				        Ok(Self { documents: records.into_records(), headers })

				    }

				}

				impl<R: Read> Iterator for CSVDocumentDeserializer<R> {

				    type Item = anyhow::Result<Object>;

				    fn next(&mut self) -> Option<Self::Item> {

				        let csv_document = self.documents.next()?;

				        match csv_document {

				            Ok(csv_document) => {

				                let mut document = Object::new();

				                for ((field_name, field_type), value) in

				                    self.headers.iter().zip(csv_document.into_iter())

				                {

				                    let parsed_value: Result<Value, ParseFloatError> = match field_type {

				                        AllowedType::Number => {

				                            value.parse::<f64>().map(Value::from).map_err(Into::into)

				                        }

				                        AllowedType::String => Ok(Value::String(value.to_string())),

				                    };

				                    match parsed_value {

				                        Ok(value) => drop(document.insert(field_name.to_string(), value)),

				                        Err(_e) => {

				                            return Some(Err(anyhow::anyhow!(

				                                "Value '{}' is not a valid number",

				                                value

				                            )))

				                        }

				                    }

				                }

				                Some(Ok(document))

				            }

				            Err(e) => Some(Err(anyhow::anyhow!("Error parsing csv document: {}", e))),

				        }

				    }

				}

									
										115

benchmarks/build.rs
									
										Normal file
									
												View File
												
				@@ -0,0 +1,115 @@

				use std::fs::File;

				use std::io::{Cursor, Read, Seek, Write};

				use std::path::{Path, PathBuf};

				use std::{env, fs};

				use bytes::Bytes;

				use convert_case::{Case, Casing};

				use flate2::read::GzDecoder;

				use reqwest::IntoUrl;

				const BASE_URL: &str = "https://milli-benchmarks.fra1.digitaloceanspaces.com/datasets";

				const DATASET_SONGS: (&str, &str) = ("smol-songs", "csv");

				const DATASET_SONGS_1_2: (&str, &str) = ("smol-songs-1_2", "csv");

				const DATASET_SONGS_3_4: (&str, &str) = ("smol-songs-3_4", "csv");

				const DATASET_SONGS_4_4: (&str, &str) = ("smol-songs-4_4", "csv");

				const DATASET_WIKI: (&str, &str) = ("smol-wiki-articles", "csv");

				const DATASET_WIKI_1_2: (&str, &str) = ("smol-wiki-articles-1_2", "csv");

				const DATASET_WIKI_3_4: (&str, &str) = ("smol-wiki-articles-3_4", "csv");

				const DATASET_WIKI_4_4: (&str, &str) = ("smol-wiki-articles-4_4", "csv");

				const DATASET_MOVIES: (&str, &str) = ("movies", "json");

				const DATASET_MOVIES_1_2: (&str, &str) = ("movies-1_2", "json");

				const DATASET_MOVIES_3_4: (&str, &str) = ("movies-3_4", "json");

				const DATASET_MOVIES_4_4: (&str, &str) = ("movies-4_4", "json");

				const DATASET_NESTED_MOVIES: (&str, &str) = ("nested_movies", "json");

				const DATASET_GEO: (&str, &str) = ("smol-all-countries", "jsonl");

				const ALL_DATASETS: &[(&str, &str)] = &[

				    DATASET_SONGS,

				    DATASET_SONGS_1_2,

				    DATASET_SONGS_3_4,

				    DATASET_SONGS_4_4,

				    DATASET_WIKI,

				    DATASET_WIKI_1_2,

				    DATASET_WIKI_3_4,

				    DATASET_WIKI_4_4,

				    DATASET_MOVIES,

				    DATASET_MOVIES_1_2,

				    DATASET_MOVIES_3_4,

				    DATASET_MOVIES_4_4,

				    DATASET_NESTED_MOVIES,

				    DATASET_GEO,

				];

				/// The name of the environment variable used to select the path

				/// of the directory containing the datasets

				const BASE_DATASETS_PATH_KEY: &str = "MILLI_BENCH_DATASETS_PATH";

				fn main() -> anyhow::Result<()> {

				    let out_dir = PathBuf::from(env::var(BASE_DATASETS_PATH_KEY).unwrap_or(env::var("OUT_DIR")?));

				    let benches_dir = PathBuf::from(env::var("CARGO_MANIFEST_DIR")?).join("benches");

				    let mut manifest_paths_file = File::create(benches_dir.join("datasets_paths.rs"))?;

				    write!(

				        manifest_paths_file,

				        r#"//! This file is generated by the build script.

				//! Do not modify by hand, use the build.rs file.

				#![allow(dead_code)]

				"#

				    )?;

				    writeln!(manifest_paths_file)?;

				    for (dataset, extension) in ALL_DATASETS {

				        let out_path = out_dir.join(dataset);

				        let out_file = out_path.with_extension(extension);

				        writeln!(

				            &mut manifest_paths_file,

				            r#"pub const {}: &str = {:?};"#,

				            dataset.to_case(Case::ScreamingSnake),

				            out_file.display(),

				        )?;

				        if out_file.exists() {

				            eprintln!(

				                "The dataset {} already exists on the file system and will not be downloaded again",

				                out_path.display(),

				            );

				            continue;

				        }

				        let url = format!("{}/{}.{}.gz", BASE_URL, dataset, extension);

				        eprintln!("downloading: {}", url);

				        let bytes = retry(|| download_dataset(url.clone()), 10)?;

				        eprintln!("{} downloaded successfully", url);

				        eprintln!("uncompressing in {}", out_file.display());

				        uncompress_in_file(bytes, &out_file)?;

				    }

				    Ok(())

				}

				fn retry<Ok, Err>(fun: impl Fn() -> Result<Ok, Err>, times: usize) -> Result<Ok, Err> {

				    for _ in 0..times {

				        if let ok @ Ok(_) = fun() {

				            return ok;

				        }

				    }

				    fun()

				}

				fn download_dataset<U: IntoUrl>(url: U) -> anyhow::Result<Cursor<Bytes>> {

				    let bytes =

				        reqwest::blocking::Client::builder().timeout(None).build()?.get(url).send()?.bytes()?;

				    Ok(Cursor::new(bytes))

				}

				fn uncompress_in_file<R: Read + Seek, P: AsRef<Path>>(bytes: R, path: P) -> anyhow::Result<()> {

				    let path = path.as_ref();

				    let mut gz = GzDecoder::new(bytes);

				    let mut dataset = Vec::new();

				    gz.read_to_end(&mut dataset)?;

				    fs::write(path, dataset)?;

				    Ok(())

				}

									
										38

benchmarks/scripts/compare.sh
									
										Executable file
									
												View File
												
				@@ -0,0 +1,38 @@

				#!/usr/bin/env bash

				# Requirements:

				# - critcmp. See: https://github.com/BurntSushi/critcmp

				# - curl

				# Usage

				# $ bash compare.sh json_file1 json_file1

				# ex: bash compare.sh songs_main_09a4321.json songs_geosearch_24ec456.json

				# Checking that critcmp is installed

				command -v critcmp > /dev/null 2>&1

				if [[ "$?" -ne 0 ]]; then

				    echo 'You must install critcmp to make this script work.'

				    echo 'See: https://github.com/BurntSushi/critcmp'

				    echo '  $ cargo install critcmp'

				    exit 1

				fi

				s3_url='https://milli-benchmarks.fra1.digitaloceanspaces.com/critcmp_results'

				for file in $@

				do

				    file_s3_url="$s3_url/$file"

				    file_local_path="/tmp/$file"

				    if [[ ! -f $file_local_path ]]; then

				        curl $file_s3_url --output $file_local_path --silent

				        if [[ "$?" -ne 0 ]]; then

				            echo 'curl command failed.'

				            exit 1

				        fi

				    fi

				done

				path_list=$(echo " $@" | sed 's/ / \/tmp\//g')

				critcmp $path_list

									
										14

benchmarks/scripts/list.sh
									
										Executable file
									
												View File
												
				@@ -0,0 +1,14 @@

				#!/usr/bin/env bash

				# Requirements:

				# - curl

				# - grep

				res=$(curl -s https://milli-benchmarks.fra1.digitaloceanspaces.com | grep -o '<Key>[^<]\+' | cut -c 5- | grep critcmp_results/ | cut -c 18-)

				for pattern in "$@"

				do

					res=$(echo "$res" | grep $pattern)

				done

				echo "$res"

									
										5

benchmarks/src/lib.rs
									
										Normal file
									
												View File
												
				@@ -0,0 +1,5 @@

				//! This library is only used to isolate the benchmarks

				//! from the original milli library.

				//!

				//! It does not include interesting functions for milli library

				//! users only for milli contributors.

									
										6

bors.toml
									
												View File
												
				@@ -1,8 +1,8 @@

				status = [

				    'Tests on ubuntu-18.04',

				    'Tests on macos-latest',

				    'Tests on windows-latest',

				    # 'Run Clippy',

				    'Tests on macos-12',

				    'Tests on windows-2022',

				    'Run Clippy',

				    'Run Rustfmt',

				    'Run tests in debug',

				]

									
										18

build-info/Cargo.toml
									
										Normal file
									
												View File
												
				@@ -0,0 +1,18 @@

				[package]

				name = "build-info"

				version.workspace = true

				authors.workspace = true

				description.workspace = true

				homepage.workspace = true

				readme.workspace = true

				edition.workspace = true

				license.workspace = true

				# See more keys and their definitions at https://doc.rust-lang.org/cargo/reference/manifest.html

				[dependencies]

				time = { version = "0.3.34", features = ["parsing"] }

				[build-dependencies]

				anyhow = "1.0.80"

				vergen-git2 = "1.0.0-beta.2"

									
										22

build-info/build.rs
									
										Normal file
									
												View File
												
				@@ -0,0 +1,22 @@

				fn main() {

				    if let Err(err) = emit_git_variables() {

				        println!("cargo:warning=vergen: {}", err);

				    }

				}

				fn emit_git_variables() -> anyhow::Result<()> {

				    // Note: any code that needs VERGEN_ environment variables should take care to define them manually in the Dockerfile and pass them

				    // in the corresponding GitHub workflow (publish_docker.yml).

				    // This is due to the Dockerfile building the binary outside of the git directory.

				    let mut builder = vergen_git2::Git2Builder::default();

				    builder.branch(true);

				    builder.commit_timestamp(true);

				    builder.commit_message(true);

				    builder.describe(true, true, None);

				    builder.sha(false);

				    let git2 = builder.build()?;

				    vergen_git2::Emitter::default().fail_on_error().add_instructions(&git2)?.emit()

				}

									
										203

build-info/src/lib.rs
									
										Normal file
									
												View File
												
				@@ -0,0 +1,203 @@

				use time::format_description::well_known::Iso8601;

				#[derive(Debug, Clone)]

				pub struct BuildInfo {

				    pub branch: Option<&'static str>,

				    pub describe: Option<DescribeResult>,

				    pub commit_sha1: Option<&'static str>,

				    pub commit_msg: Option<&'static str>,

				    pub commit_timestamp: Option<time::OffsetDateTime>,

				}

				impl BuildInfo {

				    pub fn from_build() -> Self {

				        let branch: Option<&'static str> = option_env!("VERGEN_GIT_BRANCH");

				        let describe = DescribeResult::from_build();

				        let commit_sha1 = option_env!("VERGEN_GIT_SHA");

				        let commit_msg = option_env!("VERGEN_GIT_COMMIT_MESSAGE");

				        let commit_timestamp = option_env!("VERGEN_GIT_COMMIT_TIMESTAMP");

				        let commit_timestamp = commit_timestamp.and_then(|commit_timestamp| {

				            time::OffsetDateTime::parse(commit_timestamp, &Iso8601::DEFAULT).ok()

				        });

				        Self { branch, describe, commit_sha1, commit_msg, commit_timestamp }

				    }

				}

				#[derive(Debug, Clone, Copy, PartialEq, Eq, Hash)]

				pub enum DescribeResult {

				    Prototype { name: &'static str },

				    Release { version: &'static str, major: u64, minor: u64, patch: u64 },

				    Prerelease { version: &'static str, major: u64, minor: u64, patch: u64, rc: u64 },

				    NotATag { describe: &'static str },

				}

				impl DescribeResult {

				    pub fn new(describe: &'static str) -> Self {

				        if let Some(name) = prototype_name(describe) {

				            Self::Prototype { name }

				        } else if let Some(release) = release_version(describe) {

				            release

				        } else if let Some(prerelease) = prerelease_version(describe) {

				            prerelease

				        } else {

				            Self::NotATag { describe }

				        }

				    }

				    pub fn from_build() -> Option<Self> {

				        let describe: &'static str = option_env!("VERGEN_GIT_DESCRIBE")?;

				        Some(Self::new(describe))

				    }

				    pub fn as_tag(&self) -> Option<&'static str> {

				        match self {

				            DescribeResult::Prototype { name } => Some(name),

				            DescribeResult::Release { version, .. } => Some(version),

				            DescribeResult::Prerelease { version, .. } => Some(version),

				            DescribeResult::NotATag { describe: _ } => None,

				        }

				    }

				    pub fn as_prototype(&self) -> Option<&'static str> {

				        match self {

				            DescribeResult::Prototype { name } => Some(name),

				            DescribeResult::Release { .. }

				            | DescribeResult::Prerelease { .. }

				            | DescribeResult::NotATag { .. } => None,

				        }

				    }

				}

				/// Parses the input as a prototype name.

				///

				/// Returns `Some(prototype_name)` if the following conditions are met on this value:

				///

				/// 1. starts with `prototype-`,

				/// 2. ends with `-<some_number>`,

				/// 3. does not end with `<some_number>-<some_number>`.

				///

				/// Otherwise, returns `None`.

				fn prototype_name(describe: &'static str) -> Option<&'static str> {

				    if !describe.starts_with("prototype-") {

				        return None;

				    }

				    let mut rsplit_prototype = describe.rsplit('-');

				    // last component MUST be a number

				    rsplit_prototype.next()?.parse::<u64>().ok()?;

				    // before than last component SHALL NOT be a number

				    rsplit_prototype.next()?.parse::<u64>().err()?;

				    Some(describe)

				}

				fn release_version(describe: &'static str) -> Option<DescribeResult> {

				    if !describe.starts_with('v') {

				        return None;

				    }

				    // full release version don't contain a `-`

				    if describe.contains('-') {

				        return None;

				    }

				    // full release version parse as vX.Y.Z, with X, Y, Z numbers.

				    let mut dots = describe[1..].split('.');

				    let major: u64 = dots.next()?.parse().ok()?;

				    let minor: u64 = dots.next()?.parse().ok()?;

				    let patch: u64 = dots.next()?.parse().ok()?;

				    if dots.next().is_some() {

				        return None;

				    }

				    Some(DescribeResult::Release { version: describe, major, minor, patch })

				}

				fn prerelease_version(describe: &'static str) -> Option<DescribeResult> {

				    // prerelease version is in the shape vM.N.P-rc.C

				    let mut hyphen = describe.rsplit('-');

				    let prerelease = hyphen.next()?;

				    if !prerelease.starts_with("rc.") {

				        return None;

				    }

				    let rc: u64 = prerelease[3..].parse().ok()?;

				    let release = hyphen.next()?;

				    let DescribeResult::Release { version: _, major, minor, patch } = release_version(release)?

				    else {

				        return None;

				    };

				    Some(DescribeResult::Prerelease { version: describe, major, minor, patch, rc })

				}

				#[cfg(test)]

				mod test {

				    use super::DescribeResult;

				    fn assert_not_a_tag(describe: &'static str) {

				        assert_eq!(DescribeResult::NotATag { describe }, DescribeResult::new(describe))

				    }

				    fn assert_proto(describe: &'static str) {

				        assert_eq!(DescribeResult::Prototype { name: describe }, DescribeResult::new(describe))

				    }

				    fn assert_release(describe: &'static str, major: u64, minor: u64, patch: u64) {

				        assert_eq!(

				            DescribeResult::Release { version: describe, major, minor, patch },

				            DescribeResult::new(describe)

				        )

				    }

				    fn assert_prerelease(describe: &'static str, major: u64, minor: u64, patch: u64, rc: u64) {

				        assert_eq!(

				            DescribeResult::Prerelease { version: describe, major, minor, patch, rc },

				            DescribeResult::new(describe)

				        )

				    }

				    #[test]

				    fn not_a_tag() {

				        assert_not_a_tag("whatever-fuzzy");

				        assert_not_a_tag("whatever-fuzzy-5-ggg-dirty");

				        assert_not_a_tag("whatever-fuzzy-120-ggg-dirty");

				        // technically a tag, but not a proto nor a version, so not parsed as a tag

				        assert_not_a_tag("whatever");

				        // dirty version

				        assert_not_a_tag("v1.7.0-1-ggga-dirty");

				        assert_not_a_tag("v1.7.0-rc.1-1-ggga-dirty");

				        // after version

				        assert_not_a_tag("v1.7.0-1-ggga");

				        assert_not_a_tag("v1.7.0-rc.1-1-ggga");

				        // after proto

				        assert_not_a_tag("protoype-tag-0-1-ggga");

				        assert_not_a_tag("protoype-tag-0-1-ggga-dirty");

				    }

				    #[test]

				    fn prototype() {

				        assert_proto("prototype-tag-0");

				        assert_proto("prototype-tag-10");

				        assert_proto("prototype-long-name-tag-10");

				    }

				    #[test]

				    fn release() {

				        assert_release("v1.7.2", 1, 7, 2);

				    }

				    #[test]

				    fn prerelease() {

				        assert_prerelease("v1.7.2-rc.3", 1, 7, 2, 3);

				    }

				}

									
										134

config.toml
									
										Normal file
									
												View File
												
				@@ -0,0 +1,134 @@

				# This file shows the default configuration of Meilisearch.

				# All variables are defined here: https://www.meilisearch.com/docs/learn/configuration/instance_options#environment-variables

				# Designates the location where database files will be created and retrieved.

				# https://www.meilisearch.com/docs/learn/configuration/instance_options#database-path

				db_path = "./data.ms"

				# Configures the instance's environment. Value must be either `production` or `development`.

				# https://www.meilisearch.com/docs/learn/configuration/instance_options#environment

				env = "development"

				# The address on which the HTTP server will listen.

				http_addr = "localhost:7700"

				# Sets the instance's master key, automatically protecting all routes except GET /health.

				# https://www.meilisearch.com/docs/learn/configuration/instance_options#master-key

				# master_key = "YOUR_MASTER_KEY_VALUE"

				# Deactivates Meilisearch's built-in telemetry when provided.

				# Meilisearch automatically collects data from all instances that do not opt out using this flag.

				# All gathered data is used solely for the purpose of improving Meilisearch, and can be deleted at any time.

				# https://www.meilisearch.com/docs/learn/configuration/instance_options#disable-analytics

				# no_analytics = true

				# Sets the maximum size of accepted payloads.

				# https://www.meilisearch.com/docs/learn/configuration/instance_options#payload-limit-size

				http_payload_size_limit = "100 MB"

				# Defines how much detail should be present in Meilisearch's logs.

				# Meilisearch currently supports six log levels, listed in order of increasing verbosity:  `OFF`, `ERROR`, `WARN`, `INFO`, `DEBUG`, `TRACE`

				# https://www.meilisearch.com/docs/learn/configuration/instance_options#log-level

				log_level = "INFO"

				# Sets the maximum amount of RAM Meilisearch can use when indexing.

				# https://www.meilisearch.com/docs/learn/configuration/instance_options#max-indexing-memory

				# max_indexing_memory = "2 GiB"

				# Sets the maximum number of threads Meilisearch can use during indexing.

				# https://www.meilisearch.com/docs/learn/configuration/instance_options#max-indexing-threads

				# max_indexing_threads = 4

				#############

				### DUMPS ###

				#############

				# Sets the directory where Meilisearch will create dump files.

				# https://www.meilisearch.com/docs/learn/configuration/instance_options#dump-directory

				dump_dir = "dumps/"

				# Imports the dump file located at the specified path. Path must point to a .dump file.

				# https://www.meilisearch.com/docs/learn/configuration/instance_options#import-dump

				# import_dump = "./path/to/my/file.dump"

				# Prevents Meilisearch from throwing an error when `import_dump` does not point to a valid dump file.

				# https://www.meilisearch.com/docs/learn/configuration/instance_options#ignore-missing-dump

				ignore_missing_dump = false

				# Prevents a Meilisearch instance with an existing database from throwing an error when using `import_dump`.

				# https://www.meilisearch.com/docs/learn/configuration/instance_options#ignore-dump-if-db-exists

				ignore_dump_if_db_exists = false

				#################

				### SNAPSHOTS ###

				#################

				# Enables scheduled snapshots when true, disable when false (the default).

				# If the value is given as an integer, then enables the scheduled snapshot with the passed value as the interval

				# between each snapshot, in seconds.

				# https://www.meilisearch.com/docs/learn/configuration/instance_options#schedule-snapshot-creation

				schedule_snapshot = false

				# Sets the directory where Meilisearch will store snapshots.

				# https://www.meilisearch.com/docs/learn/configuration/instance_options#snapshot-destination

				snapshot_dir = "snapshots/"

				# Launches Meilisearch after importing a previously-generated snapshot at the given filepath.

				# https://www.meilisearch.com/docs/learn/configuration/instance_options#import-snapshot

				# import_snapshot = "./path/to/my/snapshot"

				# Prevents a Meilisearch instance from throwing an error when `import_snapshot` does not point to a valid snapshot file.

				# https://www.meilisearch.com/docs/learn/configuration/instance_options#ignore-missing-snapshot

				ignore_missing_snapshot = false

				# Prevents a Meilisearch instance with an existing database from throwing an error when using `import_snapshot`.

				# https://www.meilisearch.com/docs/learn/configuration/instance_options#ignore-snapshot-if-db-exists

				ignore_snapshot_if_db_exists = false

				###########

				### SSL ###

				###########

				# Enables client authentication in the specified path.

				# https://www.meilisearch.com/docs/learn/configuration/instance_options#ssl-authentication-path

				# ssl_auth_path = "./path/to/root"

				# Sets the server's SSL certificates.

				# https://www.meilisearch.com/docs/learn/configuration/instance_options#ssl-certificates-path

				# ssl_cert_path = "./path/to/certfile"

				# Sets the server's SSL key files.

				# https://www.meilisearch.com/docs/learn/configuration/instance_options#ssl-key-path

				# ssl_key_path = "./path/to/private-key"

				# Sets the server's OCSP file.

				# https://www.meilisearch.com/docs/learn/configuration/instance_options#ssl-ocsp-path

				# ssl_ocsp_path = "./path/to/ocsp-file"

				# Makes SSL authentication mandatory.

				# https://www.meilisearch.com/docs/learn/configuration/instance_options#ssl-require-auth

				ssl_require_auth = false

				# Activates SSL session resumption.

				# https://www.meilisearch.com/docs/learn/configuration/instance_options#ssl-resumption

				ssl_resumption = false

				# Activates SSL tickets.

				# https://www.meilisearch.com/docs/learn/configuration/instance_options#ssl-tickets

				ssl_tickets = false

				#############################

				### Experimental features ###

				#############################

				# Experimental metrics feature. For more information, see: <https://github.com/meilisearch/meilisearch/discussions/3518>

				# Enables the Prometheus metrics on the `GET /metrics` endpoint.

				experimental_enable_metrics = false

				# Experimental RAM reduction during indexing, do not use in production, see: <https://github.com/meilisearch/product/discussions/652>

				experimental_reduce_indexing_memory_usage = false

				# Experimentally reduces the maximum number of tasks that will be processed at once, see: <https://github.com/orgs/meilisearch/discussions/713>

				# experimental_max_number_of_batched_tasks = 100

									
										247

download-latest.sh
									
												View File
												
				@@ -1,129 +1,51 @@

				#!/bin/sh

				# COLORS

				# This script can optionally use a GitHub token to increase your request limit (for example, if using this script in a CI).

				# To use a GitHub token, pass it through the GITHUB_PAT environment variable.

				# GLOBALS

				# Colors

				RED='\033[31m'

				GREEN='\033[32m'

				DEFAULT='\033[0m'

				# GLOBALS

				GREP_SEMVER_REGEXP='v\([0-9]*\)[.]\([0-9]*\)[.]\([0-9]*\)$' # i.e. v[number].[number].[number]

				# Project name

				PNAME='meilisearch'

				# GitHub API address

				GITHUB_API='https://api.github.com/repos/meilisearch/meilisearch/releases'

				# GitHub Release address

				GITHUB_REL='https://github.com/meilisearch/meilisearch/releases/download/'

				# FUNCTIONS

				# semverParseInto and semverLT from https://github.com/cloudflare/semver_bash/blob/master/semver.sh

				# usage: semverParseInto version major minor patch special

				# version: the string version

				# major, minor, patch, special: will be assigned by the function

				semverParseInto() {

				    local RE='[^0-9]*\([0-9]*\)[.]\([0-9]*\)[.]\([0-9]*\)\([0-9A-Za-z-]*\)'

				    #MAJOR

				    eval $2=`echo $1 | sed -e "s#$RE#\1#"`

				    #MINOR

				    eval $3=`echo $1 | sed -e "s#$RE#\2#"`

				    #PATCH

				    eval $4=`echo $1 | sed -e "s#$RE#\3#"`

				    #SPECIAL

				    eval $5=`echo $1 | sed -e "s#$RE#\4#"`

				}

				# usage: semverLT version1 version2

				semverLT() {

				    local MAJOR_A=0

				    local MINOR_A=0

				    local PATCH_A=0

				    local SPECIAL_A=0

				    local MAJOR_B=0

				    local MINOR_B=0

				    local PATCH_B=0

				    local SPECIAL_B=0

				    semverParseInto $1 MAJOR_A MINOR_A PATCH_A SPECIAL_A

				    semverParseInto $2 MAJOR_B MINOR_B PATCH_B SPECIAL_B

				    if [ $MAJOR_A -lt $MAJOR_B ]; then

				        return 0

				    fi

				    if [ $MAJOR_A -le $MAJOR_B ] && [ $MINOR_A -lt $MINOR_B ]; then

				        return 0

				    fi

				    if [ $MAJOR_A -le $MAJOR_B ] && [ $MINOR_A -le $MINOR_B ] && [ $PATCH_A -lt $PATCH_B ]; then

				        return 0

				    fi

				    if [ "_$SPECIAL_A"  == '_' ] && [ "_$SPECIAL_B"  == '_' ] ; then

				        return 1

				    fi

				    if [ "_$SPECIAL_A"  == '_' ] && [ "_$SPECIAL_B"  != '_' ] ; then

				        return 1

				    fi

				    if [ "_$SPECIAL_A"  != '_' ] && [ "_$SPECIAL_B"  == '_' ] ; then

				        return 0

				    fi

				    if [ "_$SPECIAL_A" < "_$SPECIAL_B" ]; then

				        return 0

				    fi

				    return 1

				}

				# Get a token from https://github.com/settings/tokens to increasae rate limit (from 60 to 5000), make sure the token scope is set to 'public_repo'

				# Create GITHUB_PAT enviroment variable once you aquired the token to start using it

				# Returns the tag of the latest stable release (in terms of semver and not of release date)

				# Gets the version of the latest stable version of Meilisearch by setting the $latest variable.

				# Returns 0 in case of success, 1 otherwise.

				get_latest() {

				    temp_file='temp_file' # temp_file needed because the grep would start before the download is over

				    # temp_file is needed because the grep would start before the download is over

				    temp_file=$(mktemp -q /tmp/$PNAME.XXXXXXXXX)

				    latest_release="$GITHUB_API/latest"

				    if [ $? -ne 0 ]; then

				        echo "$0: Can't create temp file."

				        fetch_release_failure_usage

				        exit 1

				    fi

				    if [ -z "$GITHUB_PAT" ]; then

				        curl -s 'https://api.github.com/repos/meilisearch/meilisearch/releases' > "$temp_file" || return 1

				        curl -s "$latest_release" > "$temp_file" || return 1

				    else

				        curl -H "Authorization: token $GITHUB_PAT" -s 'https://api.github.com/repos/meilisearch/meilisearch/releases' > "$temp_file" || return 1

				        curl -H "Authorization: token $GITHUB_PAT" -s "$latest_release" > "$temp_file" || return 1

				    fi

				    releases=$(cat "$temp_file" | \

				        grep -E '"tag_name":|"draft":|"prerelease":' \

				        | tr -d ',"' | cut -d ':' -f2 | tr -d ' ')

				        # Returns a list of [tag_name draft_boolean prerelease_boolean ...]

				        # Ex: v0.10.1 false false v0.9.1-rc.1 false true v0.9.0 false false...

				    i=0

				    latest=''

				    current_tag=''

				    for release_info in $releases; do

				        if [ $i -eq 0 ]; then # Cheking tag_name

				            if echo "$release_info" | grep -q "$GREP_SEMVER_REGEXP"; then # If it's not an alpha or beta release

				                current_tag=$release_info

				            else

				                current_tag=''

				            fi

				            i=1

				        elif [ $i -eq 1 ]; then # Checking draft boolean

				            if [ "$release_info" = 'true' ]; then

				                current_tag=''

				            fi

				            i=2

				        elif [ $i -eq 2 ]; then # Checking prerelease boolean

				            if [ "$release_info" = 'true' ]; then

				                current_tag=''

				            fi

				            i=0

				            if [ "$current_tag" != '' ]; then # If the current_tag is valid

				                if [ "$latest" = '' ]; then # If there is no latest yet

				                    latest="$current_tag"

				                else

				                    semverLT $current_tag $latest # Comparing latest and the current tag

				                    if [ $? -eq 1 ]; then

				                        latest="$current_tag"

				                    fi

				                fi

				            fi

				        fi

				    done

				    latest="$(cat "$temp_file" | grep '"tag_name":' | cut -d ':' -f2 | tr -d '"' | tr -d ',' | tr -d ' ')"

				    rm -f "$temp_file"

				    return 0

				}

				# Gets the OS by setting the $os variable

				# Gets the OS by setting the $os variable.

				# Returns 0 in case of success, 1 otherwise.

				get_os() {

				    os_name=$(uname -s)

				@@ -134,7 +56,7 @@ get_os() {

				    'Linux')

				        os='linux'

				        ;;

					'MINGW'*)

				     'MINGW'*)

				        os='windows'

				        ;;

				    *)

				@@ -143,7 +65,7 @@ get_os() {

				    return 0

				}

				# Gets the architecture by setting the $archi variable

				# Gets the architecture by setting the $archi variable.

				# Returns 0 in case of success, 1 otherwise.

				get_archi() {

				    architecture=$(uname -m)

				@@ -152,8 +74,9 @@ get_archi() {

				        archi='amd64'

				        ;;

				    'arm64')

				        if [ $os = 'macos' ]; then # MacOS M1

				            archi='amd64'

				        # macOS M1/M2

				        if [ $os = 'macos' ]; then

				            archi='apple-silicon'

				        else

				            archi='aarch64'

				        fi

				@@ -171,70 +94,74 @@ success_usage() {

				    printf "$GREEN%s\n$DEFAULT" "Meilisearch $latest binary successfully downloaded as '$binary_name' file."

				    echo ''

				    echo 'Run it:'

				    echo '    $ ./meilisearch'

				    echo "    $ ./$PNAME"

				    echo 'Usage:'

				    echo '    $ ./meilisearch --help'

				    echo "    $ ./$PNAME --help"

				}

				not_available_failure_usage() {

				    printf "$RED%s\n$DEFAULT" 'ERROR: Meilisearch binary is not available for your OS distribution or your architecture yet.'

				    echo ''

				    echo 'However, you can easily compile the binary from the source files.'

				    echo 'Follow the steps at the page ("Source" tab): https://docs.meilisearch.com/learn/getting_started/installation.html'

				    echo 'Follow the steps at the page ("Source" tab): https://www.meilisearch.com/docs/learn/getting_started/installation'

				}

				fetch_release_failure_usage() {

				    echo ''

				    printf "$RED%s\n$DEFAULT" 'ERROR: Impossible to get the latest stable version of Meilisearch.'

				    echo 'Please let us know about this issue: https://github.com/meilisearch/meilisearch/issues/new/choose'

				    echo ''

				    echo 'In the meantime, you can manually download the appropriate binary from the GitHub release assets here: https://github.com/meilisearch/meilisearch/releases/latest'

				}

				fill_release_variables() {

				    # Fill $latest variable.

				    if ! get_latest; then

				        fetch_release_failure_usage

				        exit 1

				    fi

				    if [ "$latest" = '' ]; then

				        fetch_release_failure_usage

				        exit 1

				     fi

				     # Fill $os variable.

				     if ! get_os; then

				        not_available_failure_usage

				        exit 1

				     fi

				     # Fill $archi variable.

				     if ! get_archi; then

				        not_available_failure_usage

				        exit 1

				     fi

				}

				download_binary() {

				    fill_release_variables

				    echo "Downloading Meilisearch binary $latest for $os, architecture $archi..."

				    case "$os" in

				        'windows')

				            release_file="$PNAME-$os-$archi.exe"

				            binary_name="$PNAME.exe"

				            ;;

				        *)

				            release_file="$PNAME-$os-$archi"

				            binary_name="$PNAME"

				    esac

				    # Fetch the Meilisearch binary.

				    curl --fail -OL "$GITHUB_REL/$latest/$release_file"

				    if [ $? -ne 0 ]; then

				        fetch_release_failure_usage

				        exit 1

				    fi

				    mv "$release_file" "$binary_name"

				    chmod 744 "$binary_name"

				    success_usage

				}

				# MAIN

				# Fill $latest variable

				if ! get_latest; then

				    fetch_release_failure_usage # TO CHANGE

				    exit 1

				fi

				if [ "$latest" = '' ]; then

				    fetch_release_failure_usage

				    exit 1

				fi

				# Fill $os variable

				if ! get_os; then

				    not_available_failure_usage

				    exit 1

				fi

				# Fill $archi variable

				if ! get_archi; then

				    not_available_failure_usage

				    exit 1

				fi

				echo "Downloading Meilisearch binary $latest for $os, architecture $archi..."

				case "$os" in

				    'windows')

				        release_file="meilisearch-$os-$archi.exe"

				        binary_name='meilisearch.exe'

				        ;;

				    *)

				        release_file="meilisearch-$os-$archi"

				        binary_name='meilisearch'

				esac

				# Fetch the Meilisearch binary

				link="https://github.com/meilisearch/meilisearch/releases/download/$latest/$release_file"

				curl --fail -OL "$link"

				if [ $? -ne 0 ]; then

				    fetch_release_failure_usage

				    exit 1

				fi

				mv "$release_file" "$binary_name"

				chmod 744 "$binary_name"

				success_usage

				main() {

				    download_binary

				}

				main

									
										35

dump/Cargo.toml
									
										Normal file
									
												View File
												
				@@ -0,0 +1,35 @@

				[package]

				name = "dump"

				publish = false

				version.workspace = true

				authors.workspace = true

				description.workspace = true

				edition.workspace = true

				homepage.workspace = true

				readme.workspace = true

				license.workspace = true

				[dependencies]

				anyhow = "1.0.79"

				flate2 = "1.0.28"

				http = "0.2.11"

				meilisearch-auth = { path = "../meilisearch-auth" }

				meilisearch-types = { path = "../meilisearch-types" }

				once_cell = "1.19.0"

				regex = "1.10.2"

				roaring = { version = "0.10.2", features = ["serde"] }

				serde = { version = "1.0.195", features = ["derive"] }

				serde_json = { version = "1.0.111", features = ["preserve_order"] }

				tar = "0.4.40"

				tempfile = "3.9.0"

				thiserror = "1.0.56"

				time = { version = "0.3.31", features = ["serde-well-known", "formatting", "parsing", "macros"] }

				tracing = "0.1.40"

				uuid = { version = "1.6.1", features = ["serde", "v4"] }

				[dev-dependencies]

				big_s = "1.0.2"

				maplit = "1.0.2"

				meili-snap = { path = "../meili-snap" }

				meilisearch-types = { path = "../meilisearch-types" }

									
										17

dump/README.md
									
										Normal file
									
												View File
												
				@@ -0,0 +1,17 @@

				```

				dump

				├── indexes

				│   ├── cattos

				│   │   ├── documents.jsonl

				│   │   └── settings.json

				│   └── doggos

				│       ├── documents.jsonl

				│       └── settings.json

				├── instance-uid.uuid

				├── keys.jsonl

				├── metadata.json

				└── tasks

				    ├── update_files

				    │   └── [task_id].jsonl

				    └── queue.jsonl

				```

									
										34

dump/src/error.rs
									
										Normal file
									
												View File
												
				@@ -0,0 +1,34 @@

				use meilisearch_types::error::{Code, ErrorCode};

				use thiserror::Error;

				#[derive(Debug, Error)]

				pub enum Error {

				    #[error("Bad index name.")]

				    BadIndexName,

				    #[error("Malformed task.")]

				    MalformedTask,

				    #[error(transparent)]

				    Io(#[from] std::io::Error),

				    #[error(transparent)]

				    Serde(#[from] serde_json::Error),

				    #[error(transparent)]

				    Uuid(#[from] uuid::Error),

				}

				impl ErrorCode for Error {

				    fn error_code(&self) -> Code {

				        match self {

				            Error::Io(e) => e.error_code(),

				            // These errors either happen when creating a dump and don't need any error code,

				            // or come from an internal bad deserialization.

				            Error::Serde(_) => Code::Internal,

				            Error::Uuid(_) => Code::Internal,

				            // all these errors should never be raised when creating a dump, thus no error code should be associated.

				            Error::BadIndexName => Code::Internal,

				            Error::MalformedTask => Code::Internal,

				        }

				    }

				}

									
										493

dump/src/lib.rs
									
										Normal file
									
												View File
												
				@@ -0,0 +1,493 @@

				#![allow(clippy::type_complexity)]

				#![allow(clippy::wrong_self_convention)]

				use meilisearch_types::error::ResponseError;

				use meilisearch_types::keys::Key;

				use meilisearch_types::milli::update::IndexDocumentsMethod;

				use meilisearch_types::settings::Unchecked;

				use meilisearch_types::tasks::{Details, IndexSwap, KindWithContent, Status, Task, TaskId};

				use meilisearch_types::InstanceUid;

				use roaring::RoaringBitmap;

				use serde::{Deserialize, Serialize};

				use time::OffsetDateTime;

				mod error;

				mod reader;

				mod writer;

				pub use error::Error;

				pub use reader::{DumpReader, UpdateFile};

				pub use writer::DumpWriter;

				const CURRENT_DUMP_VERSION: Version = Version::V6;

				type Result<T> = std::result::Result<T, Error>;

				#[derive(Debug, Clone, PartialEq, Eq, Serialize, Deserialize)]

				#[serde(rename_all = "camelCase")]

				pub struct Metadata {

				    pub dump_version: Version,

				    pub db_version: String,

				    #[serde(with = "time::serde::rfc3339")]

				    pub dump_date: OffsetDateTime,

				}

				#[derive(Debug, Clone, PartialEq, Eq, Serialize, Deserialize)]

				#[serde(rename_all = "camelCase")]

				pub struct IndexMetadata {

				    pub uid: String,

				    pub primary_key: Option<String>,

				    #[serde(with = "time::serde::rfc3339")]

				    pub created_at: OffsetDateTime,

				    #[serde(with = "time::serde::rfc3339")]

				    pub updated_at: OffsetDateTime,

				}

				#[derive(Debug, Clone, Copy, PartialEq, Eq, Deserialize, Serialize)]

				pub enum Version {

				    V1,

				    V2,

				    V3,

				    V4,

				    V5,

				    V6,

				}

				#[derive(Debug, PartialEq, Serialize, Deserialize)]

				#[serde(rename_all = "camelCase")]

				pub struct TaskDump {

				    pub uid: TaskId,

				    #[serde(default)]

				    pub index_uid: Option<String>,

				    pub status: Status,

				    #[serde(rename = "type")]

				    pub kind: KindDump,

				    #[serde(skip_serializing_if = "Option::is_none")]

				    pub canceled_by: Option<TaskId>,

				    #[serde(skip_serializing_if = "Option::is_none")]

				    pub details: Option<Details>,

				    #[serde(skip_serializing_if = "Option::is_none")]

				    pub error: Option<ResponseError>,

				    #[serde(with = "time::serde::rfc3339")]

				    pub enqueued_at: OffsetDateTime,

				    #[serde(

				        with = "time::serde::rfc3339::option",

				        skip_serializing_if = "Option::is_none",

				        default

				    )]

				    pub started_at: Option<OffsetDateTime>,

				    #[serde(

				        with = "time::serde::rfc3339::option",

				        skip_serializing_if = "Option::is_none",

				        default

				    )]

				    pub finished_at: Option<OffsetDateTime>,

				}

				// A `Kind` specific version made for the dump. If modified you may break the dump.

				#[derive(Debug, PartialEq, Serialize, Deserialize)]

				#[serde(rename_all = "camelCase")]

				pub enum KindDump {

				    DocumentImport {

				        primary_key: Option<String>,

				        method: IndexDocumentsMethod,

				        documents_count: u64,

				        allow_index_creation: bool,

				    },

				    DocumentDeletion {

				        documents_ids: Vec<String>,

				    },

				    DocumentClear,

				    DocumentDeletionByFilter {

				        filter: serde_json::Value,

				    },

				    Settings {

				        settings: Box<meilisearch_types::settings::Settings<Unchecked>>,

				        is_deletion: bool,

				        allow_index_creation: bool,

				    },

				    IndexDeletion,

				    IndexCreation {

				        primary_key: Option<String>,

				    },

				    IndexUpdate {

				        primary_key: Option<String>,

				    },

				    IndexSwap {

				        swaps: Vec<IndexSwap>,

				    },

				    TaskCancelation {

				        query: String,

				        tasks: RoaringBitmap,

				    },

				    TasksDeletion {

				        query: String,

				        tasks: RoaringBitmap,

				    },

				    DumpCreation {

				        keys: Vec<Key>,

				        instance_uid: Option<InstanceUid>,

				    },

				    SnapshotCreation,

				}

				impl From<Task> for TaskDump {

				    fn from(task: Task) -> Self {

				        TaskDump {

				            uid: task.uid,

				            index_uid: task.index_uid().map(|uid| uid.to_string()),

				            status: task.status,

				            kind: task.kind.into(),

				            canceled_by: task.canceled_by,

				            details: task.details,

				            error: task.error,

				            enqueued_at: task.enqueued_at,

				            started_at: task.started_at,

				            finished_at: task.finished_at,

				        }

				    }

				}

				impl From<KindWithContent> for KindDump {

				    fn from(kind: KindWithContent) -> Self {

				        match kind {

				            KindWithContent::DocumentAdditionOrUpdate {

				                primary_key,

				                method,

				                documents_count,

				                allow_index_creation,

				                ..

				            } => KindDump::DocumentImport {

				                primary_key,

				                method,

				                documents_count,

				                allow_index_creation,

				            },

				            KindWithContent::DocumentDeletion { documents_ids, .. } => {

				                KindDump::DocumentDeletion { documents_ids }

				            }

				            KindWithContent::DocumentDeletionByFilter { filter_expr, .. } => {

				                KindDump::DocumentDeletionByFilter { filter: filter_expr }

				            }

				            KindWithContent::DocumentClear { .. } => KindDump::DocumentClear,

				            KindWithContent::SettingsUpdate {

				                new_settings,

				                is_deletion,

				                allow_index_creation,

				                ..

				            } => KindDump::Settings { settings: new_settings, is_deletion, allow_index_creation },

				            KindWithContent::IndexDeletion { .. } => KindDump::IndexDeletion,

				            KindWithContent::IndexCreation { primary_key, .. } => {

				                KindDump::IndexCreation { primary_key }

				            }

				            KindWithContent::IndexUpdate { primary_key, .. } => {

				                KindDump::IndexUpdate { primary_key }

				            }

				            KindWithContent::IndexSwap { swaps } => KindDump::IndexSwap { swaps },

				            KindWithContent::TaskCancelation { query, tasks } => {

				                KindDump::TaskCancelation { query, tasks }

				            }

				            KindWithContent::TaskDeletion { query, tasks } => {

				                KindDump::TasksDeletion { query, tasks }

				            }

				            KindWithContent::DumpCreation { keys, instance_uid } => {

				                KindDump::DumpCreation { keys, instance_uid }

				            }

				            KindWithContent::SnapshotCreation => KindDump::SnapshotCreation,

				        }

				    }

				}

				#[cfg(test)]

				pub(crate) mod test {

				    use std::fs::File;

				    use std::io::Seek;

				    use std::str::FromStr;

				    use big_s::S;

				    use maplit::{btreemap, btreeset};

				    use meilisearch_types::facet_values_sort::FacetValuesSort;

				    use meilisearch_types::features::RuntimeTogglableFeatures;

				    use meilisearch_types::index_uid_pattern::IndexUidPattern;

				    use meilisearch_types::keys::{Action, Key};

				    use meilisearch_types::milli;

				    use meilisearch_types::milli::update::Setting;

				    use meilisearch_types::settings::{Checked, FacetingSettings, Settings};

				    use meilisearch_types::tasks::{Details, Status};

				    use serde_json::{json, Map, Value};

				    use time::macros::datetime;

				    use uuid::Uuid;

				    use crate::reader::Document;

				    use crate::{DumpReader, DumpWriter, IndexMetadata, KindDump, TaskDump, Version};

				    pub fn create_test_instance_uid() -> Uuid {

				        Uuid::parse_str("9e15e977-f2ae-4761-943f-1eaf75fd736d").unwrap()

				    }

				    pub fn create_test_index_metadata() -> IndexMetadata {

				        IndexMetadata {

				            uid: S("doggo"),

				            primary_key: None,

				            created_at: datetime!(2022-11-20 12:00 UTC),

				            updated_at: datetime!(2022-11-21 00:00 UTC),

				        }

				    }

				    pub fn create_test_documents() -> Vec<Map<String, Value>> {

				        vec![

				            json!({ "id": 1, "race": "golden retriever", "name": "paul", "age": 4 })

				                .as_object()

				                .unwrap()

				                .clone(),

				            json!({ "id": 2, "race": "bernese mountain", "name": "tamo", "age": 6 })

				                .as_object()

				                .unwrap()

				                .clone(),

				            json!({ "id": 3, "race": "great pyrenees", "name": "patou", "age": 5 })

				                .as_object()

				                .unwrap()

				                .clone(),

				        ]

				    }

				    pub fn create_test_settings() -> Settings<Checked> {

				        let settings = Settings {

				            displayed_attributes: Setting::Set(vec![S("race"), S("name")]),

				            searchable_attributes: Setting::Set(vec![S("name"), S("race")]),

				            filterable_attributes: Setting::Set(btreeset! { S("race"), S("age") }),

				            sortable_attributes: Setting::Set(btreeset! { S("age") }),

				            ranking_rules: Setting::NotSet,

				            stop_words: Setting::NotSet,

				            non_separator_tokens: Setting::NotSet,

				            separator_tokens: Setting::NotSet,

				            dictionary: Setting::NotSet,

				            synonyms: Setting::NotSet,

				            distinct_attribute: Setting::NotSet,

				            proximity_precision: Setting::NotSet,

				            typo_tolerance: Setting::NotSet,

				            faceting: Setting::Set(FacetingSettings {

				                max_values_per_facet: Setting::Set(111),

				                sort_facet_values_by: Setting::Set(

				                    btreemap! { S("age") => FacetValuesSort::Count },

				                ),

				            }),

				            pagination: Setting::NotSet,

				            embedders: Setting::NotSet,

				            _kind: std::marker::PhantomData,

				        };

				        settings.check()

				    }

				    pub fn create_test_tasks() -> Vec<(TaskDump, Option<Vec<Document>>)> {

				        vec![

				            (

				                TaskDump {

				                    uid: 0,

				                    index_uid: Some(S("doggo")),

				                    status: Status::Succeeded,

				                    kind: KindDump::DocumentImport {

				                        method: milli::update::IndexDocumentsMethod::UpdateDocuments,

				                        allow_index_creation: true,

				                        primary_key: Some(S("bone")),

				                        documents_count: 12,

				                    },

				                    canceled_by: None,

				                    details: Some(Details::DocumentAdditionOrUpdate {

				                        received_documents: 12,

				                        indexed_documents: Some(10),

				                    }),

				                    error: None,

				                    enqueued_at: datetime!(2022-11-11 0:00 UTC),

				                    started_at: Some(datetime!(2022-11-20 0:00 UTC)),

				                    finished_at: Some(datetime!(2022-11-21 0:00 UTC)),

				                },

				                None,

				            ),

				            (

				                TaskDump {

				                    uid: 1,

				                    index_uid: Some(S("doggo")),

				                    status: Status::Enqueued,

				                    kind: KindDump::DocumentImport {

				                        method: milli::update::IndexDocumentsMethod::UpdateDocuments,

				                        allow_index_creation: true,

				                        primary_key: None,

				                        documents_count: 2,

				                    },

				                    canceled_by: None,

				                    details: Some(Details::DocumentAdditionOrUpdate {

				                        received_documents: 2,

				                        indexed_documents: None,

				                    }),

				                    error: None,

				                    enqueued_at: datetime!(2022-11-11 0:00 UTC),

				                    started_at: None,

				                    finished_at: None,

				                },

				                Some(vec![

				                    json!({ "id": 4, "race": "leonberg" }).as_object().unwrap().clone(),

				                    json!({ "id": 5, "race": "patou" }).as_object().unwrap().clone(),

				                ]),

				            ),

				            (

				                TaskDump {

				                    uid: 5,

				                    index_uid: Some(S("catto")),

				                    status: Status::Enqueued,

				                    kind: KindDump::IndexDeletion,

				                    canceled_by: None,

				                    details: None,

				                    error: None,

				                    enqueued_at: datetime!(2022-11-15 0:00 UTC),

				                    started_at: None,

				                    finished_at: None,

				                },

				                None,

				            ),

				        ]

				    }

				    pub fn create_test_api_keys() -> Vec<Key> {

				        vec![

				            Key {

				                description: Some(S("The main key to manage all the doggos")),

				                name: Some(S("doggos_key")),

				                uid: Uuid::from_str("9f8a34da-b6b2-42f0-939b-dbd4c3448655").unwrap(),

				                actions: vec![Action::DocumentsAll],

				                indexes: vec![IndexUidPattern::from_str("doggos").unwrap()],

				                expires_at: Some(datetime!(4130-03-14 12:21 UTC)),

				                created_at: datetime!(1960-11-15 0:00 UTC),

				                updated_at: datetime!(2022-11-10 0:00 UTC),

				            },

				            Key {

				                description: Some(S("The master key for everything and even the doggos")),

				                name: Some(S("master_key")),

				                uid: Uuid::from_str("4622f717-1c00-47bb-a494-39d76a49b591").unwrap(),

				                actions: vec![Action::All],

				                indexes: vec![IndexUidPattern::all()],

				                expires_at: None,

				                created_at: datetime!(0000-01-01 00:01 UTC),

				                updated_at: datetime!(1964-05-04 17:25 UTC),

				            },

				            Key {

				                description: Some(S("The useless key to for nothing nor the doggos")),

				                name: Some(S("useless_key")),

				                uid: Uuid::from_str("fb80b58b-0a34-412f-8ba7-1ce868f8ac5c").unwrap(),

				                actions: vec![],

				                indexes: vec![],

				                expires_at: None,

				                created_at: datetime!(400-02-29 0:00 UTC),

				                updated_at: datetime!(1024-02-29 0:00 UTC),

				            },

				        ]

				    }

				    pub fn create_test_dump() -> File {

				        let instance_uid = create_test_instance_uid();

				        let dump = DumpWriter::new(Some(instance_uid)).unwrap();

				        // ========== Adding an index

				        let documents = create_test_documents();

				        let settings = create_test_settings();

				        let mut index = dump.create_index("doggos", &create_test_index_metadata()).unwrap();

				        for document in &documents {

				            index.push_document(document).unwrap();

				        }

				        index.flush().unwrap();

				        index.settings(&settings).unwrap();

				        // ========== pushing the task queue

				        let tasks = create_test_tasks();

				        let mut task_queue = dump.create_tasks_queue().unwrap();

				        for (task, update_file) in &tasks {

				            let mut update = task_queue.push_task(task).unwrap();

				            if let Some(update_file) = update_file {

				                for u in update_file {

				                    update.push_document(u).unwrap();

				                }

				            }

				        }

				        task_queue.flush().unwrap();

				        // ========== pushing the api keys

				        let api_keys = create_test_api_keys();

				        let mut keys = dump.create_keys().unwrap();

				        for key in &api_keys {

				            keys.push_key(key).unwrap();

				        }

				        keys.flush().unwrap();

				        // ========== experimental features

				        let features = create_test_features();

				        dump.create_experimental_features(features).unwrap();

				        // create the dump

				        let mut file = tempfile::tempfile().unwrap();

				        dump.persist_to(&mut file).unwrap();

				        file.rewind().unwrap();

				        file

				    }

				    fn create_test_features() -> RuntimeTogglableFeatures {

				        RuntimeTogglableFeatures { vector_store: true, ..Default::default() }

				    }

				    #[test]

				    fn test_creating_and_read_dump() {

				        let mut file = create_test_dump();

				        let mut dump = DumpReader::open(&mut file).unwrap();

				        // ==== checking the top level infos

				        assert_eq!(dump.version(), Version::V6);

				        assert!(dump.date().is_some());

				        assert_eq!(dump.instance_uid().unwrap().unwrap(), create_test_instance_uid());

				        // ==== checking the index

				        let mut indexes = dump.indexes().unwrap();

				        let mut index = indexes.next().unwrap().unwrap();

				        assert!(indexes.next().is_none()); // there was only one index in the dump

				        for (document, expected) in index.documents().unwrap().zip(create_test_documents()) {

				            assert_eq!(document.unwrap(), expected);

				        }

				        assert_eq!(index.settings().unwrap(), create_test_settings());

				        assert_eq!(index.metadata(), &create_test_index_metadata());

				        drop(index);

				        drop(indexes);

				        // ==== checking the task queue

				        for (task, expected) in dump.tasks().unwrap().zip(create_test_tasks()) {

				            let (task, content_file) = task.unwrap();

				            assert_eq!(task, expected.0);

				            if let Some(expected_update) = expected.1 {

				                assert!(

				                    content_file.is_some(),

				                    "A content file was expected for the task {}.",

				                    expected.0.uid

				                );

				                let updates = content_file.unwrap().collect::<Result<Vec<_>, _>>().unwrap();

				                assert_eq!(updates, expected_update);

				            }

				        }

				        // ==== checking the keys

				        for (key, expected) in dump.keys().unwrap().zip(create_test_api_keys()) {

				            assert_eq!(key.unwrap(), expected);

				        }

				        // ==== checking the features

				        let expected = create_test_features();

				        assert_eq!(dump.features().unwrap().unwrap(), expected);

				    }

				}

									
										5

dump/src/reader/compat/mod.rs
									
										Normal file
									
												View File
												
				@@ -0,0 +1,5 @@

				pub mod v1_to_v2;

				pub mod v2_to_v3;

				pub mod v3_to_v4;

				pub mod v4_to_v5;

				pub mod v5_to_v6;

38

dump/src/reader/compat/snapshots/dumpreadercompat__v1_to_v2testcompat_v1_v2-3.snap Normal file

View File

@@ -0,0 +1,38 @@
 ---
 source: dump/src/reader/compat/v1_to_v2.rs
 expression: products.settings().unwrap()
 ---
 {
   "displayedAttributes": [
     "*"
   ],
   "searchableAttributes": [
     "*"
   ],
   "filterableAttributes": [],
   "sortableAttributes": [],
   "rankingRules": [
     "typo",
     "words",
     "proximity",
     "attribute",
     "exactness"
   ],
   "stopWords": [],
   "synonyms": {
     "android": [
       "phone",
       "smartphone"
     ],
     "iphone": [
       "phone",
       "smartphone"
     ],
     "phone": [
       "android",
       "iphone",
       "smartphone"
     ]
   },
   "distinctAttribute": null
 }

31

dump/src/reader/compat/snapshots/dumpreadercompat__v1_to_v2testcompat_v1_v2-6.snap Normal file

View File

@@ -0,0 +1,31 @@
 ---
 source: dump/src/reader/compat/v1_to_v2.rs
 expression: movies.settings().unwrap()
 ---
 {
   "displayedAttributes": [
     "*"
   ],
   "searchableAttributes": [
     "*"
   ],
   "filterableAttributes": [
     "genres",
     "id"
   ],
   "sortableAttributes": [
     "genres",
     "id"
   ],
   "rankingRules": [
     "typo",
     "words",
     "proximity",
     "attribute",
     "exactness",
     "release_date:asc"
   ],
   "stopWords": [],
   "synonyms": {},
   "distinctAttribute": null
 }

24

dump/src/reader/compat/snapshots/dumpreadercompat__v1_to_v2testcompat_v1_v2-9.snap Normal file

View File

@@ -0,0 +1,24 @@
 ---
 source: dump/src/reader/compat/v1_to_v2.rs
 expression: spells.settings().unwrap()
 ---
 {
   "displayedAttributes": [
     "*"
   ],
   "searchableAttributes": [
     "*"
   ],
   "filterableAttributes": [],
   "sortableAttributes": [],
   "rankingRules": [
     "typo",
     "words",
     "proximity",
     "attribute",
     "exactness"
   ],
   "stopWords": [],
   "synonyms": {},
   "distinctAttribute": null
 }

23

dump/src/reader/compat/snapshots/dumpreadercompat__v2_to_v3testcompat_v2_v3-11.snap Normal file

View File

@@ -0,0 +1,23 @@
 ---
 source: dump/src/reader/compat/v2_to_v3.rs
 expression: movies2.settings().unwrap()
 ---
 {
   "displayedAttributes": [
     "*"
   ],
   "searchableAttributes": [
     "*"
   ],
   "filterableAttributes": [],
   "rankingRules": [
     "words",
     "typo",
     "proximity",
     "attribute",
     "exactness"
   ],
   "stopWords": [],
   "synonyms": {},
   "distinctAttribute": null
 }

23

dump/src/reader/compat/snapshots/dumpreadercompat__v2_to_v3testcompat_v2_v3-14.snap Normal file

View File

@@ -0,0 +1,23 @@
 ---
 source: dump/src/reader/compat/v2_to_v3.rs
 expression: spells.settings().unwrap()
 ---
 {
   "displayedAttributes": [
     "*"
   ],
   "searchableAttributes": [
     "*"
   ],
   "filterableAttributes": [],
   "rankingRules": [
     "words",
     "typo",
     "proximity",
     "attribute",
     "exactness"
   ],
   "stopWords": [],
   "synonyms": {},
   "distinctAttribute": null
 }

37

dump/src/reader/compat/snapshots/dumpreadercompat__v2_to_v3testcompat_v2_v3-5.snap Normal file

View File

@@ -0,0 +1,37 @@
 ---
 source: dump/src/reader/compat/v2_to_v3.rs
 expression: products.settings().unwrap()
 ---
 {
   "displayedAttributes": [
     "*"
   ],
   "searchableAttributes": [
     "*"
   ],
   "filterableAttributes": [],
   "rankingRules": [
     "words",
     "typo",
     "proximity",
     "attribute",
     "exactness"
   ],
   "stopWords": [],
   "synonyms": {
     "android": [
       "phone",
       "smartphone"
     ],
     "iphone": [
       "phone",
       "smartphone"
     ],
     "phone": [
       "android",
       "iphone",
       "smartphone"
     ]
   },
   "distinctAttribute": null
 }

24

dump/src/reader/compat/snapshots/dumpreadercompat__v2_to_v3testcompat_v2_v3-8.snap Normal file

View File

@@ -0,0 +1,24 @@
 ---
 source: dump/src/reader/compat/v2_to_v3.rs
 expression: movies.settings().unwrap()
 ---
 {
   "displayedAttributes": [
     "*"
   ],
   "searchableAttributes": [
     "*"
   ],
   "filterableAttributes": [],
   "rankingRules": [
     "words",
     "typo",
     "proximity",
     "attribute",
     "exactness",
     "release_date:asc"
   ],
   "stopWords": [],
   "synonyms": {},
   "distinctAttribute": null
 }

25

dump/src/reader/compat/snapshots/dumpreadercompat__v3_to_v4testcompat_v3_v4-12.snap Normal file

View File

@@ -0,0 +1,25 @@
 ---
 source: dump/src/reader/compat/v3_to_v4.rs
 expression: movies2.settings().unwrap()
 ---
 {
   "displayedAttributes": [
     "*"
   ],
   "searchableAttributes": [
     "*"
   ],
   "filterableAttributes": [],
   "sortableAttributes": [],
   "rankingRules": [
     "words",
     "typo",
     "proximity",
     "attribute",
     "sort",
     "exactness"
   ],
   "stopWords": [],
   "synonyms": {},
   "distinctAttribute": null
 }

25

dump/src/reader/compat/snapshots/dumpreadercompat__v3_to_v4testcompat_v3_v4-15.snap Normal file

View File

@@ -0,0 +1,25 @@
 ---
 source: dump/src/reader/compat/v3_to_v4.rs
 expression: spells.settings().unwrap()
 ---
 {
   "displayedAttributes": [
     "*"
   ],
   "searchableAttributes": [
     "*"
   ],
   "filterableAttributes": [],
   "sortableAttributes": [],
   "rankingRules": [
     "words",
     "typo",
     "proximity",
     "attribute",
     "sort",
     "exactness"
   ],
   "stopWords": [],
   "synonyms": {},
   "distinctAttribute": null
 }

39

dump/src/reader/compat/snapshots/dumpreadercompat__v3_to_v4testcompat_v3_v4-6.snap Normal file

View File

@@ -0,0 +1,39 @@
 ---
 source: dump/src/reader/compat/v3_to_v4.rs
 expression: products.settings().unwrap()
 ---
 {
   "displayedAttributes": [
     "*"
   ],
   "searchableAttributes": [
     "*"
   ],
   "filterableAttributes": [],
   "sortableAttributes": [],
   "rankingRules": [
     "words",
     "typo",
     "proximity",
     "attribute",
     "sort",
     "exactness"
   ],
   "stopWords": [],
   "synonyms": {
     "android": [
       "phone",
       "smartphone"
     ],
     "iphone": [
       "phone",
       "smartphone"
     ],
     "phone": [
       "android",
       "iphone",
       "smartphone"
     ]
   },
   "distinctAttribute": null
 }

31

dump/src/reader/compat/snapshots/dumpreadercompat__v3_to_v4testcompat_v3_v4-9.snap Normal file

View File

@@ -0,0 +1,31 @@
 ---
 source: dump/src/reader/compat/v3_to_v4.rs
 expression: movies.settings().unwrap()
 ---
 {
   "displayedAttributes": [
     "*"
   ],
   "searchableAttributes": [
     "*"
   ],
   "filterableAttributes": [
     "genres",
     "id"
   ],
   "sortableAttributes": [
     "release_date"
   ],
   "rankingRules": [
     "words",
     "typo",
     "proximity",
     "attribute",
     "sort",
     "exactness",
     "release_date:asc"
   ],
   "stopWords": [],
   "synonyms": {},
   "distinctAttribute": null
 }

56

dump/src/reader/compat/snapshots/dumpreadercompat__v4_to_v5testcompat_v4_v5-12.snap Normal file

View File

@@ -0,0 +1,56 @@
 ---
 source: dump/src/reader/compat/v4_to_v5.rs
 expression: spells.settings().unwrap()
 ---
 {
   "displayedAttributes": "Reset",
   "searchableAttributes": "Reset",
   "filterableAttributes": {
     "Set": []
   },
   "sortableAttributes": {
     "Set": []
   },
   "rankingRules": {
     "Set": [
       "words",
       "typo",
       "proximity",
       "attribute",
       "sort",
       "exactness"
     ]
   },
   "stopWords": {
     "Set": []
   },
   "synonyms": {
     "Set": {}
   },
   "distinctAttribute": "Reset",
   "typoTolerance": {
     "Set": {
       "enabled": {
         "Set": true
       },
       "minWordSizeForTypos": {
         "Set": {
           "oneTypo": {
             "Set": 5
           },
           "twoTypos": {
             "Set": 9
           }
         }
       },
       "disableOnWords": {
         "Set": []
       },
       "disableOnAttributes": {
         "Set": []
       }
     }
   },
   "faceting": "NotSet",
   "pagination": "NotSet"
 }

70

dump/src/reader/compat/snapshots/dumpreadercompat__v4_to_v5testcompat_v4_v5-6.snap Normal file

View File

@@ -0,0 +1,70 @@
 ---
 source: dump/src/reader/compat/v4_to_v5.rs
 expression: products.settings().unwrap()
 ---
 {
   "displayedAttributes": "Reset",
   "searchableAttributes": "Reset",
   "filterableAttributes": {
     "Set": []
   },
   "sortableAttributes": {
     "Set": []
   },
   "rankingRules": {
     "Set": [
       "words",
       "typo",
       "proximity",
       "attribute",
       "sort",
       "exactness"
     ]
   },
   "stopWords": {
     "Set": []
   },
   "synonyms": {
     "Set": {
       "android": [
         "phone",
         "smartphone"
       ],
       "iphone": [
         "phone",
         "smartphone"
       ],
       "phone": [
         "android",
         "iphone",
         "smartphone"
       ]
     }
   },
   "distinctAttribute": "Reset",
   "typoTolerance": {
     "Set": {
       "enabled": {
         "Set": true
       },
       "minWordSizeForTypos": {
         "Set": {
           "oneTypo": {
             "Set": 5
           },
           "twoTypos": {
             "Set": 9
           }
         }
       },
       "disableOnWords": {
         "Set": []
       },
       "disableOnAttributes": {
         "Set": []
       }
     }
   },
   "faceting": "NotSet",
   "pagination": "NotSet"
 }

62

dump/src/reader/compat/snapshots/dumpreadercompat__v4_to_v5testcompat_v4_v5-9.snap Normal file

View File

@@ -0,0 +1,62 @@
 ---
 source: dump/src/reader/compat/v4_to_v5.rs
 expression: movies.settings().unwrap()
 ---
 {
   "displayedAttributes": "Reset",
   "searchableAttributes": "Reset",
   "filterableAttributes": {
     "Set": [
       "genres",
       "id"
     ]
   },
   "sortableAttributes": {
     "Set": [
       "release_date"
     ]
   },
   "rankingRules": {
     "Set": [
       "words",
       "typo",
       "proximity",
       "attribute",
       "sort",
       "exactness",
       "release_date:asc"
     ]
   },
   "stopWords": {
     "Set": []
   },
   "synonyms": {
     "Set": {}
   },
   "distinctAttribute": "Reset",
   "typoTolerance": {
     "Set": {
       "enabled": {
         "Set": true
       },
       "minWordSizeForTypos": {
         "Set": {
           "oneTypo": {
             "Set": 5
           },
           "twoTypos": {
             "Set": 9
           }
         }
       },
       "disableOnWords": {
         "Set": []
       },
       "disableOnAttributes": {
         "Set": []
       }
     }
   },
   "faceting": "NotSet",
   "pagination": "NotSet"
 }

40

dump/src/reader/compat/snapshots/dumpreadercompat__v5_to_v6testcompat_v5_v6-12.snap Normal file

View File

@@ -0,0 +1,40 @@
 ---
 source: dump/src/reader/compat/v5_to_v6.rs
 expression: spells.settings().unwrap()
 ---
 {
   "displayedAttributes": [
     "*"
   ],
   "searchableAttributes": [
     "*"
   ],
   "filterableAttributes": [],
   "sortableAttributes": [],
   "rankingRules": [
     "words",
     "typo",
     "proximity",
     "attribute",
     "sort",
     "exactness"
   ],
   "stopWords": [],
   "synonyms": {},
   "distinctAttribute": null,
   "typoTolerance": {
     "enabled": true,
     "minWordSizeForTypos": {
       "oneTypo": 5,
       "twoTypos": 9
     },
     "disableOnWords": [],
     "disableOnAttributes": []
   },
   "faceting": {
     "maxValuesPerFacet": 100
   },
   "pagination": {
     "maxTotalHits": 1000
   }
 }

54

dump/src/reader/compat/snapshots/dumpreadercompat__v5_to_v6testcompat_v5_v6-6.snap Normal file

View File

@@ -0,0 +1,54 @@
 ---
 source: dump/src/reader/compat/v5_to_v6.rs
 expression: products.settings().unwrap()
 ---
 {
   "displayedAttributes": [
     "*"
   ],
   "searchableAttributes": [
     "*"
   ],
   "filterableAttributes": [],
   "sortableAttributes": [],
   "rankingRules": [
     "words",
     "typo",
     "proximity",
     "attribute",
     "sort",
     "exactness"
   ],
   "stopWords": [],
   "synonyms": {
     "android": [
       "phone",
       "smartphone"
     ],
     "iphone": [
       "phone",
       "smartphone"
     ],
     "phone": [
       "android",
       "iphone",
       "smartphone"
     ]
   },
   "distinctAttribute": null,
   "typoTolerance": {
     "enabled": true,
     "minWordSizeForTypos": {
       "oneTypo": 5,
       "twoTypos": 9
     },
     "disableOnWords": [],
     "disableOnAttributes": []
   },
   "faceting": {
     "maxValuesPerFacet": 100
   },
   "pagination": {
     "maxTotalHits": 1000
   }
 }

46

dump/src/reader/compat/snapshots/dumpreadercompat__v5_to_v6testcompat_v5_v6-9.snap Normal file

View File

@@ -0,0 +1,46 @@
 ---
 source: dump/src/reader/compat/v5_to_v6.rs
 expression: movies.settings().unwrap()
 ---
 {
   "displayedAttributes": [
     "*"
   ],
   "searchableAttributes": [
     "*"
   ],
   "filterableAttributes": [
     "genres",
     "id"
   ],
   "sortableAttributes": [
     "release_date"
   ],
   "rankingRules": [
     "words",
     "typo",
     "proximity",
     "attribute",
     "sort",
     "exactness",
     "release_date:asc"
   ],
   "stopWords": [],
   "synonyms": {},
   "distinctAttribute": null,
   "typoTolerance": {
     "enabled": true,
     "minWordSizeForTypos": {
       "oneTypo": 5,
       "twoTypos": 9
     },
     "disableOnWords": [],
     "disableOnAttributes": []
   },
   "faceting": {
     "maxValuesPerFacet": 100
   },
   "pagination": {
     "maxTotalHits": 1000
   }
 }

									
										410

dump/src/reader/compat/v1_to_v2.rs
									
										Normal file
									
												View File
												
				@@ -0,0 +1,410 @@

				use std::str::FromStr;

				use super::v2_to_v3::CompatV2ToV3;

				use crate::reader::{v1, v2, Document};

				use crate::Result;

				pub struct CompatV1ToV2 {

				    pub from: v1::V1Reader,

				}

				impl CompatV1ToV2 {

				    pub fn new(v1: v1::V1Reader) -> Self {

				        Self { from: v1 }

				    }

				    pub fn to_v3(self) -> CompatV2ToV3 {

				        CompatV2ToV3::Compat(self)

				    }

				    pub fn version(&self) -> crate::Version {

				        self.from.version()

				    }

				    pub fn date(&self) -> Option<time::OffsetDateTime> {

				        self.from.date()

				    }

				    pub fn index_uuid(&self) -> Vec<v2::meta::IndexUuid> {

				        self.from

				            .index_uuid()

				            .into_iter()

				            .enumerate()

				            // we use the index of the index 😬 as UUID for the index, so that we can link the v2::Task to their index

				            .map(|(index, index_uuid)| v2::meta::IndexUuid {

				                uid: index_uuid.uid,

				                uuid: uuid::Uuid::from_u128(index as u128),

				            })

				            .collect()

				    }

				    pub fn indexes(&self) -> Result<impl Iterator<Item = Result<CompatIndexV1ToV2>> + '_> {

				        Ok(self.from.indexes()?.map(|index_reader| Ok(CompatIndexV1ToV2 { from: index_reader? })))

				    }

				    pub fn tasks(

				        &mut self,

				    ) -> Box<dyn Iterator<Item = Result<(v2::Task, Option<v2::UpdateFile>)>> + '_> {

				        // Convert an error here to an iterator yielding the error

				        let indexes = match self.from.indexes() {

				            Ok(indexes) => indexes,

				            Err(err) => return Box::new(std::iter::once(Err(err))),

				        };

				        let it = indexes.enumerate().flat_map(

				            move |(index, index_reader)| -> Box<dyn Iterator<Item = _>> {

				                let index_reader = match index_reader {

				                    Ok(index_reader) => index_reader,

				                    Err(err) => return Box::new(std::iter::once(Err(err))),

				                };

				                Box::new(

				                    index_reader

				                        .tasks()

				                        // Filter out the UpdateStatus::Customs variant that is not supported in v2

				                        // and enqueued tasks, that don't contain the necessary update file in v1

				                        .filter_map(move |task| -> Option<_> {

				                            let task = match task {

				                                Ok(task) => task,

				                                Err(err) => return Some(Err(err)),

				                            };

				                            Some(Ok((

				                                v2::Task {

				                                    uuid: uuid::Uuid::from_u128(index as u128),

				                                    update: Option::from(task)?,

				                                },

				                                None,

				                            )))

				                        }),

				                )

				            },

				        );

				        Box::new(it)

				    }

				}

				pub struct CompatIndexV1ToV2 {

				    pub from: v1::V1IndexReader,

				}

				impl CompatIndexV1ToV2 {

				    pub fn metadata(&self) -> &crate::IndexMetadata {

				        self.from.metadata()

				    }

				    pub fn documents(&mut self) -> Result<Box<dyn Iterator<Item = Result<Document>> + '_>> {

				        self.from.documents().map(|it| Box::new(it) as Box<dyn Iterator<Item = _>>)

				    }

				    pub fn settings(&mut self) -> Result<v2::settings::Settings<v2::settings::Checked>> {

				        Ok(v2::settings::Settings::<v2::settings::Unchecked>::from(self.from.settings()?).check())

				    }

				}

				impl From<v1::settings::Settings> for v2::Settings<v2::Unchecked> {

				    fn from(source: v1::settings::Settings) -> Self {

				        Self {

				            displayed_attributes: option_to_setting(source.displayed_attributes)

				                .map(|displayed| displayed.into_iter().collect()),

				            searchable_attributes: option_to_setting(source.searchable_attributes),

				            filterable_attributes: option_to_setting(source.attributes_for_faceting.clone())

				                .map(|filterable| filterable.into_iter().collect()),

				            sortable_attributes: option_to_setting(source.attributes_for_faceting)

				                .map(|sortable| sortable.into_iter().collect()),

				            ranking_rules: option_to_setting(source.ranking_rules).map(|ranking_rules| {

				                ranking_rules

				                    .into_iter()

				                    .filter_map(|ranking_rule| {

				                        match v1::settings::RankingRule::from_str(&ranking_rule) {

				                            Ok(ranking_rule) => {

				                                let criterion: Option<v2::settings::Criterion> =

				                                    ranking_rule.into();

				                                criterion.as_ref().map(ToString::to_string)

				                            }

				                            Err(()) => {

				                                tracing::warn!(

				                                    "Could not import the following ranking rule: `{}`.",

				                                    ranking_rule

				                                );

				                                None

				                            }

				                        }

				                    })

				                    .collect()

				            }),

				            stop_words: option_to_setting(source.stop_words),

				            synonyms: option_to_setting(source.synonyms),

				            distinct_attribute: option_to_setting(source.distinct_attribute),

				            _kind: std::marker::PhantomData,

				        }

				    }

				}

				fn option_to_setting<T>(opt: Option<Option<T>>) -> v2::Setting<T> {

				    match opt {

				        Some(Some(t)) => v2::Setting::Set(t),

				        None => v2::Setting::NotSet,

				        Some(None) => v2::Setting::Reset,

				    }

				}

				impl From<v1::update::UpdateStatus> for Option<v2::updates::UpdateStatus> {

				    fn from(source: v1::update::UpdateStatus) -> Self {

				        use v1::update::UpdateStatus as UpdateStatusV1;

				        use v2::updates::UpdateStatus as UpdateStatusV2;

				        Some(match source {

				            UpdateStatusV1::Enqueued { content } => {

				                tracing::warn!(

				                    "Cannot import task {} (importing enqueued tasks from v1 dumps is unsupported)",

				                    content.update_id

				                );

				                tracing::warn!("Task will be skipped in the queue of imported tasks.");

				                return None;

				            }

				            UpdateStatusV1::Failed { content } => UpdateStatusV2::Failed(v2::updates::Failed {

				                from: v2::updates::Processing {

				                    from: v2::updates::Enqueued {

				                        update_id: content.update_id,

				                        meta: Option::from(content.update_type)?,

				                        enqueued_at: content.enqueued_at,

				                        content: None,

				                    },

				                    started_processing_at: content.processed_at

				                        - std::time::Duration::from_secs_f64(content.duration),

				                },

				                error: v2::ResponseError {

				                    // error code is ignored by serialization, and so always default in deserialized v2 dumps

				                    // that's a good thing, because we don't have them in v1 dump 😅

				                    code: http::StatusCode::default(),

				                    message: content.error.unwrap_or_default(),

				                    // error codes are unchanged between v1 and v2

				                    error_code: content.error_code.unwrap_or_default(),

				                    // error types are unchanged between v1 and v2

				                    error_type: content.error_type.unwrap_or_default(),

				                    // error links are unchanged between v1 and v2

				                    error_link: content.error_link.unwrap_or_default(),

				                },

				                failed_at: content.processed_at,

				            }),

				            UpdateStatusV1::Processed { content } => {

				                UpdateStatusV2::Processed(v2::updates::Processed {

				                    success: match &content.update_type {

				                        v1::update::UpdateType::ClearAll => {

				                            v2::updates::UpdateResult::DocumentDeletion { deleted: u64::MAX }

				                        }

				                        v1::update::UpdateType::Customs => v2::updates::UpdateResult::Other,

				                        v1::update::UpdateType::DocumentsAddition { number } => {

				                            v2::updates::UpdateResult::DocumentsAddition(

				                                v2::updates::DocumentAdditionResult { nb_documents: *number },

				                            )

				                        }

				                        v1::update::UpdateType::DocumentsPartial { number } => {

				                            v2::updates::UpdateResult::DocumentsAddition(

				                                v2::updates::DocumentAdditionResult { nb_documents: *number },

				                            )

				                        }

				                        v1::update::UpdateType::DocumentsDeletion { number } => {

				                            v2::updates::UpdateResult::DocumentDeletion { deleted: *number as u64 }

				                        }

				                        v1::update::UpdateType::Settings { .. } => v2::updates::UpdateResult::Other,

				                    },

				                    processed_at: content.processed_at,

				                    from: v2::updates::Processing {

				                        from: v2::updates::Enqueued {

				                            update_id: content.update_id,

				                            meta: Option::from(content.update_type)?,

				                            enqueued_at: content.enqueued_at,

				                            content: None,

				                        },

				                        started_processing_at: content.processed_at

				                            - std::time::Duration::from_secs_f64(content.duration),

				                    },

				                })

				            }

				        })

				    }

				}

				impl From<v1::update::UpdateType> for Option<v2::updates::UpdateMeta> {

				    fn from(source: v1::update::UpdateType) -> Self {

				        Some(match source {

				            v1::update::UpdateType::ClearAll => v2::updates::UpdateMeta::ClearDocuments,

				            v1::update::UpdateType::Customs => {

				                tracing::warn!("Ignoring task with type 'Customs' that is no longer supported");

				                return None;

				            }

				            v1::update::UpdateType::DocumentsAddition { .. } => {

				                v2::updates::UpdateMeta::DocumentsAddition {

				                    method: v2::updates::IndexDocumentsMethod::ReplaceDocuments,

				                    format: v2::updates::UpdateFormat::Json,

				                    primary_key: None,

				                }

				            }

				            v1::update::UpdateType::DocumentsPartial { .. } => {

				                v2::updates::UpdateMeta::DocumentsAddition {

				                    method: v2::updates::IndexDocumentsMethod::UpdateDocuments,

				                    format: v2::updates::UpdateFormat::Json,

				                    primary_key: None,

				                }

				            }

				            v1::update::UpdateType::DocumentsDeletion { .. } => {

				                v2::updates::UpdateMeta::DeleteDocuments { ids: vec![] }

				            }

				            v1::update::UpdateType::Settings { settings } => {

				                v2::updates::UpdateMeta::Settings((*settings).into())

				            }

				        })

				    }

				}

				impl From<v1::settings::SettingsUpdate> for v2::Settings<v2::Unchecked> {

				    fn from(source: v1::settings::SettingsUpdate) -> Self {

				        let ranking_rules = v2::Setting::from(source.ranking_rules);

				        // go from the concrete types of v1 (RankingRule) to the concrete type of v2 (Criterion),

				        // and then back to string as this is what the settings manipulate

				        let ranking_rules = ranking_rules.map(|ranking_rules| {

				            ranking_rules

				                .into_iter()

				                // filter out the WordsPosition ranking rule that exists in v1 but not v2

				                .filter_map(Option::<v2::settings::Criterion>::from)

				                .map(|criterion| criterion.to_string())

				                .collect()

				        });

				        Self {

				            displayed_attributes: v2::Setting::from(source.displayed_attributes)

				                .map(|displayed_attributes| displayed_attributes.into_iter().collect()),

				            searchable_attributes: source.searchable_attributes.into(),

				            filterable_attributes: v2::Setting::from(source.attributes_for_faceting.clone())

				                .map(|attributes_for_faceting| attributes_for_faceting.into_iter().collect()),

				            sortable_attributes: v2::Setting::from(source.attributes_for_faceting)

				                .map(|attributes_for_faceting| attributes_for_faceting.into_iter().collect()),

				            ranking_rules,

				            stop_words: source.stop_words.into(),

				            synonyms: source.synonyms.into(),

				            distinct_attribute: source.distinct_attribute.into(),

				            _kind: std::marker::PhantomData,

				        }

				    }

				}

				impl From<v1::settings::RankingRule> for Option<v2::settings::Criterion> {

				    fn from(source: v1::settings::RankingRule) -> Self {

				        match source {

				            v1::settings::RankingRule::Typo => Some(v2::settings::Criterion::Typo),

				            v1::settings::RankingRule::Words => Some(v2::settings::Criterion::Words),

				            v1::settings::RankingRule::Proximity => Some(v2::settings::Criterion::Proximity),

				            v1::settings::RankingRule::Attribute => Some(v2::settings::Criterion::Attribute),

				            v1::settings::RankingRule::WordsPosition => {

				                tracing::warn!("Removing the 'WordsPosition' ranking rule that is no longer supported, please check the resulting ranking rules of your indexes");

				                None

				            }

				            v1::settings::RankingRule::Exactness => Some(v2::settings::Criterion::Exactness),

				            v1::settings::RankingRule::Asc(field_name) => {

				                Some(v2::settings::Criterion::Asc(field_name))

				            }

				            v1::settings::RankingRule::Desc(field_name) => {

				                Some(v2::settings::Criterion::Desc(field_name))

				            }

				        }

				    }

				}

				impl<T> From<v1::settings::UpdateState<T>> for v2::Setting<T> {

				    fn from(source: v1::settings::UpdateState<T>) -> Self {

				        match source {

				            v1::settings::UpdateState::Update(new_value) => v2::Setting::Set(new_value),

				            v1::settings::UpdateState::Clear => v2::Setting::Reset,

				            v1::settings::UpdateState::Nothing => v2::Setting::NotSet,

				        }

				    }

				}

				#[cfg(test)]

				pub(crate) mod test {

				    use std::fs::File;

				    use std::io::BufReader;

				    use flate2::bufread::GzDecoder;

				    use meili_snap::insta;

				    use tempfile::TempDir;

				    use super::*;

				    #[test]

				    fn compat_v1_v2() {

				        let dump = File::open("tests/assets/v1.dump").unwrap();

				        let dir = TempDir::new().unwrap();

				        let mut dump = BufReader::new(dump);

				        let gz = GzDecoder::new(&mut dump);

				        let mut archive = tar::Archive::new(gz);

				        archive.unpack(dir.path()).unwrap();

				        let mut dump = v1::V1Reader::open(dir).unwrap().to_v2();

				        // top level infos

				        assert_eq!(dump.date(), None);

				        // tasks

				        let tasks = dump.tasks().collect::<Result<Vec<_>>>().unwrap();

				        let (tasks, update_files): (Vec<_>, Vec<_>) = tasks.into_iter().unzip();

				        meili_snap::snapshot_hash!(meili_snap::json_string!(tasks), @"2298010973ee98cf4670787314176a3a");

				        assert_eq!(update_files.len(), 9);

				        assert!(update_files[..].iter().all(|u| u.is_none())); // no update file in dumps v1

				        // indexes

				        let mut indexes = dump.indexes().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        // the index are not ordered in any way by default

				        indexes.sort_by_key(|index| index.metadata().uid.to_string());

				        let mut products = indexes.pop().unwrap();

				        let mut movies = indexes.pop().unwrap();

				        let mut spells = indexes.pop().unwrap();

				        assert!(indexes.is_empty());

				        // products

				        insta::assert_json_snapshot!(products.metadata(), @r###"

				        {

				          "uid": "products",

				          "primaryKey": "sku",

				          "createdAt": "2022-10-02T13:23:39.976870431Z",

				          "updatedAt": "2022-10-02T13:27:54.353262482Z"

				        }

				        "###);

				        insta::assert_json_snapshot!(products.settings().unwrap());

				        let documents = products.documents().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        assert_eq!(documents.len(), 10);

				        meili_snap::snapshot_hash!(format!("{:#?}", documents), @"b01c8371aea4c7171af0d4d846a2bdca");

				        // movies

				        insta::assert_json_snapshot!(movies.metadata(), @r###"

				        {

				          "uid": "movies",

				          "primaryKey": "id",

				          "createdAt": "2022-10-02T13:15:29.477512777Z",

				          "updatedAt": "2022-10-02T13:21:12.671204856Z"

				        }

				        "###);

				        insta::assert_json_snapshot!(movies.settings().unwrap());

				        let documents = movies.documents().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        assert_eq!(documents.len(), 10);

				        meili_snap::snapshot_hash!(format!("{:#?}", documents), @"b63dbed5bbc059f3e32bc471ae699bf5");

				        // spells

				        insta::assert_json_snapshot!(spells.metadata(), @r###"

				        {

				          "uid": "dnd_spells",

				          "primaryKey": "index",

				          "createdAt": "2022-10-02T13:38:26.358882984Z",

				          "updatedAt": "2022-10-02T13:38:26.385609433Z"

				        }

				        "###);

				        insta::assert_json_snapshot!(spells.settings().unwrap());

				        let documents = spells.documents().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        assert_eq!(documents.len(), 10);

				        meili_snap::snapshot_hash!(format!("{:#?}", documents), @"aa24c0cfc733d66c396237ad44263bed");

				    }

				}

									
										512

dump/src/reader/compat/v2_to_v3.rs
									
										Normal file
									
												View File
												
				@@ -0,0 +1,512 @@

				use std::str::FromStr;

				use time::OffsetDateTime;

				use uuid::Uuid;

				use super::v1_to_v2::{CompatIndexV1ToV2, CompatV1ToV2};

				use super::v3_to_v4::CompatV3ToV4;

				use crate::reader::{v2, v3, Document};

				use crate::Result;

				pub enum CompatV2ToV3 {

				    V2(v2::V2Reader),

				    Compat(CompatV1ToV2),

				}

				impl CompatV2ToV3 {

				    pub fn new(v2: v2::V2Reader) -> CompatV2ToV3 {

				        CompatV2ToV3::V2(v2)

				    }

				    pub fn index_uuid(&self) -> Vec<v3::meta::IndexUuid> {

				        let v2_uuids = match self {

				            CompatV2ToV3::V2(from) => from.index_uuid(),

				            CompatV2ToV3::Compat(compat) => compat.index_uuid(),

				        };

				        v2_uuids

				            .into_iter()

				            .map(|index| v3::meta::IndexUuid { uid: index.uid, uuid: index.uuid })

				            .collect()

				    }

				    pub fn to_v4(self) -> CompatV3ToV4 {

				        CompatV3ToV4::Compat(self)

				    }

				    pub fn version(&self) -> crate::Version {

				        match self {

				            CompatV2ToV3::V2(from) => from.version(),

				            CompatV2ToV3::Compat(compat) => compat.version(),

				        }

				    }

				    pub fn date(&self) -> Option<time::OffsetDateTime> {

				        match self {

				            CompatV2ToV3::V2(from) => from.date(),

				            CompatV2ToV3::Compat(compat) => compat.date(),

				        }

				    }

				    pub fn instance_uid(&self) -> Result<Option<uuid::Uuid>> {

				        Ok(None)

				    }

				    pub fn indexes(&self) -> Result<impl Iterator<Item = Result<CompatIndexV2ToV3>> + '_> {

				        Ok(match self {

				            CompatV2ToV3::V2(from) => Box::new(from.indexes()?.map(|index_reader| -> Result<_> {

				                let compat = CompatIndexV2ToV3::new(index_reader?);

				                Ok(compat)

				            }))

				                as Box<dyn Iterator<Item = Result<CompatIndexV2ToV3>> + '_>,

				            CompatV2ToV3::Compat(compat) => Box::new(compat.indexes()?.map(|index_reader| {

				                let compat = CompatIndexV2ToV3::Compat(Box::new(index_reader?));

				                Ok(compat)

				            }))

				                as Box<dyn Iterator<Item = Result<CompatIndexV2ToV3>> + '_>,

				        })

				    }

				    pub fn tasks(

				        &mut self,

				    ) -> Box<

				        dyn Iterator<Item = Result<(v3::Task, Option<Box<dyn Iterator<Item = Result<Document>>>>)>>

				            + '_,

				    > {

				        let tasks = match self {

				            CompatV2ToV3::V2(from) => from.tasks(),

				            CompatV2ToV3::Compat(compat) => compat.tasks(),

				        };

				        Box::new(

				            tasks

				                .map(move |task| {

				                    task.map(|(task, content_file)| {

				                        let task = v3::Task { uuid: task.uuid, update: task.update.into() };

				                        Some((

				                            task,

				                            content_file.map(|content_file| {

				                                Box::new(content_file) as Box<dyn Iterator<Item = Result<Document>>>

				                            }),

				                        ))

				                    })

				                })

				                .filter_map(|res| res.transpose()),

				        )

				    }

				}

				pub enum CompatIndexV2ToV3 {

				    V2(v2::V2IndexReader),

				    Compat(Box<CompatIndexV1ToV2>),

				}

				impl CompatIndexV2ToV3 {

				    pub fn new(v2: v2::V2IndexReader) -> CompatIndexV2ToV3 {

				        CompatIndexV2ToV3::V2(v2)

				    }

				    pub fn metadata(&self) -> &crate::IndexMetadata {

				        match self {

				            CompatIndexV2ToV3::V2(from) => from.metadata(),

				            CompatIndexV2ToV3::Compat(compat) => compat.metadata(),

				        }

				    }

				    pub fn documents(&mut self) -> Result<Box<dyn Iterator<Item = Result<Document>> + '_>> {

				        match self {

				            CompatIndexV2ToV3::V2(from) => from

				                .documents()

				                .map(|iter| Box::new(iter) as Box<dyn Iterator<Item = Result<Document>> + '_>),

				            CompatIndexV2ToV3::Compat(compat) => compat.documents(),

				        }

				    }

				    pub fn settings(&mut self) -> Result<v3::Settings<v3::Checked>> {

				        let settings = match self {

				            CompatIndexV2ToV3::V2(from) => from.settings()?,

				            CompatIndexV2ToV3::Compat(compat) => compat.settings()?,

				        };

				        Ok(v3::Settings::<v3::Unchecked>::from(settings).check())

				    }

				}

				impl From<v2::updates::UpdateStatus> for v3::updates::UpdateStatus {

				    fn from(update: v2::updates::UpdateStatus) -> Self {

				        match update {

				            v2::updates::UpdateStatus::Processing(processing) => {

				                match (processing.from.meta.clone(), processing.from.content).try_into() {

				                    Ok(meta) => v3::updates::UpdateStatus::Processing(v3::updates::Processing {

				                        from: v3::updates::Enqueued {

				                            update_id: processing.from.update_id,

				                            meta,

				                            enqueued_at: processing.from.enqueued_at,

				                        },

				                        started_processing_at: processing.started_processing_at,

				                    }),

				                    Err(e) => {

				                        tracing::warn!("Error with task {}: {}", processing.from.update_id, e);

				                        tracing::warn!("Task will be marked as `Failed`.");

				                        v3::updates::UpdateStatus::Failed(v3::updates::Failed {

				                            from: v3::updates::Processing {

				                                from: v3::updates::Enqueued {

				                                    update_id: processing.from.update_id,

				                                    meta: update_from_unchecked_update_meta(processing.from.meta),

				                                    enqueued_at: processing.from.enqueued_at,

				                                },

				                                started_processing_at: processing.started_processing_at,

				                            },

				                            msg: e.to_string(),

				                            code: v3::Code::MalformedDump,

				                            failed_at: OffsetDateTime::now_utc(),

				                        })

				                    }

				                }

				            }

				            v2::updates::UpdateStatus::Enqueued(enqueued) => {

				                match (enqueued.meta.clone(), enqueued.content).try_into() {

				                    Ok(meta) => v3::updates::UpdateStatus::Enqueued(v3::updates::Enqueued {

				                        update_id: enqueued.update_id,

				                        meta,

				                        enqueued_at: enqueued.enqueued_at,

				                    }),

				                    Err(e) => {

				                        tracing::warn!("Error with task {}: {}", enqueued.update_id, e);

				                        tracing::warn!("Task will be marked as `Failed`.");

				                        v3::updates::UpdateStatus::Failed(v3::updates::Failed {

				                            from: v3::updates::Processing {

				                                from: v3::updates::Enqueued {

				                                    update_id: enqueued.update_id,

				                                    meta: update_from_unchecked_update_meta(enqueued.meta),

				                                    enqueued_at: enqueued.enqueued_at,

				                                },

				                                started_processing_at: OffsetDateTime::now_utc(),

				                            },

				                            msg: e.to_string(),

				                            code: v3::Code::MalformedDump,

				                            failed_at: OffsetDateTime::now_utc(),

				                        })

				                    }

				                }

				            }

				            v2::updates::UpdateStatus::Processed(processed) => {

				                v3::updates::UpdateStatus::Processed(v3::updates::Processed {

				                    success: processed.success.into(),

				                    processed_at: processed.processed_at,

				                    from: v3::updates::Processing {

				                        from: v3::updates::Enqueued {

				                            update_id: processed.from.from.update_id,

				                            // since we're never going to read the content_file again it's ok to generate a fake one.

				                            meta: update_from_unchecked_update_meta(processed.from.from.meta),

				                            enqueued_at: processed.from.from.enqueued_at,

				                        },

				                        started_processing_at: processed.from.started_processing_at,

				                    },

				                })

				            }

				            v2::updates::UpdateStatus::Aborted(aborted) => {

				                v3::updates::UpdateStatus::Aborted(v3::updates::Aborted {

				                    from: v3::updates::Enqueued {

				                        update_id: aborted.from.update_id,

				                        // since we're never going to read the content_file again it's ok to generate a fake one.

				                        meta: update_from_unchecked_update_meta(aborted.from.meta),

				                        enqueued_at: aborted.from.enqueued_at,

				                    },

				                    aborted_at: aborted.aborted_at,

				                })

				            }

				            v2::updates::UpdateStatus::Failed(failed) => {

				                v3::updates::UpdateStatus::Failed(v3::updates::Failed {

				                    from: v3::updates::Processing {

				                        from: v3::updates::Enqueued {

				                            update_id: failed.from.from.update_id,

				                            // since we're never going to read the content_file again it's ok to generate a fake one.

				                            meta: update_from_unchecked_update_meta(failed.from.from.meta),

				                            enqueued_at: failed.from.from.enqueued_at,

				                        },

				                        started_processing_at: failed.from.started_processing_at,

				                    },

				                    msg: failed.error.message,

				                    code: failed.error.error_code.into(),

				                    failed_at: failed.failed_at,

				                })

				            }

				        }

				    }

				}

				impl TryFrom<(v2::updates::UpdateMeta, Option<Uuid>)> for v3::updates::Update {

				    type Error = crate::Error;

				    fn try_from((update, uuid): (v2::updates::UpdateMeta, Option<Uuid>)) -> Result<Self> {

				        Ok(match update {

				            v2::updates::UpdateMeta::DocumentsAddition { method, format: _, primary_key }

				                if uuid.is_some() =>

				            {

				                v3::updates::Update::DocumentAddition {

				                    primary_key,

				                    method: match method {

				                        v2::updates::IndexDocumentsMethod::ReplaceDocuments => {

				                            v3::updates::IndexDocumentsMethod::ReplaceDocuments

				                        }

				                        v2::updates::IndexDocumentsMethod::UpdateDocuments => {

				                            v3::updates::IndexDocumentsMethod::UpdateDocuments

				                        }

				                    },

				                    content_uuid: uuid.unwrap(),

				                }

				            }

				            v2::updates::UpdateMeta::DocumentsAddition { .. } => {

				                return Err(crate::Error::MalformedTask)

				            }

				            v2::updates::UpdateMeta::ClearDocuments => v3::updates::Update::ClearDocuments,

				            v2::updates::UpdateMeta::DeleteDocuments { ids } => {

				                v3::updates::Update::DeleteDocuments(ids)

				            }

				            v2::updates::UpdateMeta::Settings(settings) => {

				                v3::updates::Update::Settings(settings.into())

				            }

				        })

				    }

				}

				pub fn update_from_unchecked_update_meta(update: v2::updates::UpdateMeta) -> v3::updates::Update {

				    match update {

				        v2::updates::UpdateMeta::DocumentsAddition { method, format: _, primary_key } => {

				            v3::updates::Update::DocumentAddition {

				                primary_key,

				                method: match method {

				                    v2::updates::IndexDocumentsMethod::ReplaceDocuments => {

				                        v3::updates::IndexDocumentsMethod::ReplaceDocuments

				                    }

				                    v2::updates::IndexDocumentsMethod::UpdateDocuments => {

				                        v3::updates::IndexDocumentsMethod::UpdateDocuments

				                    }

				                },

				                // we use this special uuid so we can recognize it if one day there is a bug related to this field.

				                content_uuid: Uuid::from_str("00112233-4455-6677-8899-aabbccddeeff").unwrap(),

				            }

				        }

				        v2::updates::UpdateMeta::ClearDocuments => v3::updates::Update::ClearDocuments,

				        v2::updates::UpdateMeta::DeleteDocuments { ids } => {

				            v3::updates::Update::DeleteDocuments(ids)

				        }

				        v2::updates::UpdateMeta::Settings(settings) => {

				            v3::updates::Update::Settings(settings.into())

				        }

				    }

				}

				impl From<v2::updates::UpdateResult> for v3::updates::UpdateResult {

				    fn from(result: v2::updates::UpdateResult) -> Self {

				        match result {

				            v2::updates::UpdateResult::DocumentsAddition(addition) => {

				                v3::updates::UpdateResult::DocumentsAddition(v3::updates::DocumentAdditionResult {

				                    nb_documents: addition.nb_documents,

				                })

				            }

				            v2::updates::UpdateResult::DocumentDeletion { deleted } => {

				                v3::updates::UpdateResult::DocumentDeletion { deleted }

				            }

				            v2::updates::UpdateResult::Other => v3::updates::UpdateResult::Other,

				        }

				    }

				}

				impl From<String> for v3::Code {

				    fn from(code: String) -> Self {

				        match code.as_ref() {

				            "create_index" => v3::Code::CreateIndex,

				            "index_already_exists" => v3::Code::IndexAlreadyExists,

				            "index_not_found" => v3::Code::IndexNotFound,

				            "invalid_index_uid" => v3::Code::InvalidIndexUid,

				            "invalid_state" => v3::Code::InvalidState,

				            "missing_primary_key" => v3::Code::MissingPrimaryKey,

				            "primary_key_already_present" => v3::Code::PrimaryKeyAlreadyPresent,

				            "max_fields_limit_exceeded" => v3::Code::MaxFieldsLimitExceeded,

				            "missing_document_id" => v3::Code::MissingDocumentId,

				            "invalid_document_id" => v3::Code::InvalidDocumentId,

				            "filter" => v3::Code::Filter,

				            "sort" => v3::Code::Sort,

				            "bad_parameter" => v3::Code::BadParameter,

				            "bad_request" => v3::Code::BadRequest,

				            "database_size_limit_reached" => v3::Code::DatabaseSizeLimitReached,

				            "document_not_found" => v3::Code::DocumentNotFound,

				            "internal" => v3::Code::Internal,

				            "invalid_geo_field" => v3::Code::InvalidGeoField,

				            "invalid_ranking_rule" => v3::Code::InvalidRankingRule,

				            "invalid_store" => v3::Code::InvalidStore,

				            "invalid_token" => v3::Code::InvalidToken,

				            "missing_authorization_header" => v3::Code::MissingAuthorizationHeader,

				            "no_space_left_on_device" => v3::Code::NoSpaceLeftOnDevice,

				            "dump_not_found" => v3::Code::DumpNotFound,

				            "task_not_found" => v3::Code::TaskNotFound,

				            "payload_too_large" => v3::Code::PayloadTooLarge,

				            "retrieve_document" => v3::Code::RetrieveDocument,

				            "search_documents" => v3::Code::SearchDocuments,

				            "unsupported_media_type" => v3::Code::UnsupportedMediaType,

				            "dump_already_in_progress" => v3::Code::DumpAlreadyInProgress,

				            "dump_process_failed" => v3::Code::DumpProcessFailed,

				            "invalid_content_type" => v3::Code::InvalidContentType,

				            "missing_content_type" => v3::Code::MissingContentType,

				            "malformed_payload" => v3::Code::MalformedPayload,

				            "missing_payload" => v3::Code::MissingPayload,

				            other => {

				                tracing::warn!("Unknown error code {}", other);

				                v3::Code::UnretrievableErrorCode

				            }

				        }

				    }

				}

				impl<A> From<v2::Setting<A>> for v3::Setting<A> {

				    fn from(setting: v2::Setting<A>) -> Self {

				        match setting {

				            v2::settings::Setting::Set(a) => v3::settings::Setting::Set(a),

				            v2::settings::Setting::Reset => v3::settings::Setting::Reset,

				            v2::settings::Setting::NotSet => v3::settings::Setting::NotSet,

				        }

				    }

				}

				impl<T> From<v2::Settings<T>> for v3::Settings<v3::Unchecked> {

				    fn from(settings: v2::Settings<T>) -> Self {

				        v3::Settings {

				            displayed_attributes: settings.displayed_attributes.into(),

				            searchable_attributes: settings.searchable_attributes.into(),

				            filterable_attributes: settings.filterable_attributes.into(),

				            sortable_attributes: settings.sortable_attributes.into(),

				            ranking_rules: v3::Setting::from(settings.ranking_rules).map(|criteria| {

				                criteria.into_iter().map(|criterion| patch_ranking_rules(&criterion)).collect()

				            }),

				            stop_words: settings.stop_words.into(),

				            synonyms: settings.synonyms.into(),

				            distinct_attribute: settings.distinct_attribute.into(),

				            _kind: std::marker::PhantomData,

				        }

				    }

				}

				fn patch_ranking_rules(ranking_rule: &str) -> String {

				    match v2::settings::Criterion::from_str(ranking_rule) {

				        Ok(v2::settings::Criterion::Words) => String::from("words"),

				        Ok(v2::settings::Criterion::Typo) => String::from("typo"),

				        Ok(v2::settings::Criterion::Proximity) => String::from("proximity"),

				        Ok(v2::settings::Criterion::Attribute) => String::from("attribute"),

				        Ok(v2::settings::Criterion::Sort) => String::from("sort"),

				        Ok(v2::settings::Criterion::Exactness) => String::from("exactness"),

				        Ok(v2::settings::Criterion::Asc(name)) => format!("{name}:asc"),

				        Ok(v2::settings::Criterion::Desc(name)) => format!("{name}:desc"),

				        // we want to forward the error to the current version of meilisearch

				        Err(_) => ranking_rule.to_string(),

				    }

				}

				#[cfg(test)]

				pub(crate) mod test {

				    use std::fs::File;

				    use std::io::BufReader;

				    use flate2::bufread::GzDecoder;

				    use meili_snap::insta;

				    use tempfile::TempDir;

				    use super::*;

				    #[test]

				    fn compat_v2_v3() {

				        let dump = File::open("tests/assets/v2.dump").unwrap();

				        let dir = TempDir::new().unwrap();

				        let mut dump = BufReader::new(dump);

				        let gz = GzDecoder::new(&mut dump);

				        let mut archive = tar::Archive::new(gz);

				        archive.unpack(dir.path()).unwrap();

				        let mut dump = v2::V2Reader::open(dir).unwrap().to_v3();

				        // top level infos

				        insta::assert_display_snapshot!(dump.date().unwrap(), @"2022-10-09 20:27:59.904096267 +00:00:00");

				        // tasks

				        let tasks = dump.tasks().collect::<Result<Vec<_>>>().unwrap();

				        let (tasks, mut update_files): (Vec<_>, Vec<_>) = tasks.into_iter().unzip();

				        meili_snap::snapshot_hash!(meili_snap::json_string!(tasks), @"9507711db47c7171c79bc6d57d0bed79");

				        assert_eq!(update_files.len(), 9);

				        assert!(update_files[0].is_some()); // the enqueued document addition

				        assert!(update_files[1..].iter().all(|u| u.is_none())); // everything already processed

				        let update_file = update_files.remove(0).unwrap().collect::<Result<Vec<_>>>().unwrap();

				        meili_snap::snapshot_hash!(meili_snap::json_string!(update_file), @"7b8889539b669c7b9ddba448bafa385d");

				        // indexes

				        let mut indexes = dump.indexes().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        // the index are not ordered in any way by default

				        indexes.sort_by_key(|index| index.metadata().uid.to_string());

				        let mut products = indexes.pop().unwrap();

				        let mut movies2 = indexes.pop().unwrap();

				        let mut movies = indexes.pop().unwrap();

				        let mut spells = indexes.pop().unwrap();

				        assert!(indexes.is_empty());

				        // products

				        insta::assert_json_snapshot!(products.metadata(), { ".createdAt" => "[now]", ".updatedAt" => "[now]" }, @r###"

				        {

				          "uid": "products",

				          "primaryKey": "sku",

				          "createdAt": "[now]",

				          "updatedAt": "[now]"

				        }

				        "###);

				        insta::assert_json_snapshot!(products.settings().unwrap());

				        let documents = products.documents().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        assert_eq!(documents.len(), 10);

				        meili_snap::snapshot_hash!(format!("{:#?}", documents), @"548284a84de510f71e88e6cdea495cf5");

				        // movies

				        insta::assert_json_snapshot!(movies.metadata(), { ".createdAt" => "[now]", ".updatedAt" => "[now]" }, @r###"

				        {

				          "uid": "movies",

				          "primaryKey": "id",

				          "createdAt": "[now]",

				          "updatedAt": "[now]"

				        }

				        "###);

				        insta::assert_json_snapshot!(movies.settings().unwrap());

				        let documents = movies.documents().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        assert_eq!(documents.len(), 110);

				        meili_snap::snapshot_hash!(format!("{:#?}", documents), @"d153b5a81d8b3cdcbe1dec270b574022");

				        // movies2

				        insta::assert_json_snapshot!(movies2.metadata(), { ".createdAt" => "[now]", ".updatedAt" => "[now]" }, @r###"

				        {

				          "uid": "movies_2",

				          "primaryKey": null,

				          "createdAt": "[now]",

				          "updatedAt": "[now]"

				        }

				        "###);

				        insta::assert_json_snapshot!(movies2.settings().unwrap());

				        let documents = movies2.documents().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        assert_eq!(documents.len(), 0);

				        meili_snap::snapshot_hash!(format!("{:#?}", documents), @"d751713988987e9331980363e24189ce");

				        // spells

				        insta::assert_json_snapshot!(spells.metadata(), { ".createdAt" => "[now]", ".updatedAt" => "[now]" }, @r###"

				        {

				          "uid": "dnd_spells",

				          "primaryKey": "index",

				          "createdAt": "[now]",

				          "updatedAt": "[now]"

				        }

				        "###);

				        insta::assert_json_snapshot!(spells.settings().unwrap());

				        let documents = spells.documents().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        assert_eq!(documents.len(), 10);

				        meili_snap::snapshot_hash!(format!("{:#?}", documents), @"235016433dd04262c7f2da01d1e808ce");

				    }

				}

									
										449

dump/src/reader/compat/v3_to_v4.rs
									
										Normal file
									
												View File
												
				@@ -0,0 +1,449 @@

				use super::v2_to_v3::{CompatIndexV2ToV3, CompatV2ToV3};

				use super::v4_to_v5::CompatV4ToV5;

				use crate::reader::{v3, v4, UpdateFile};

				use crate::Result;

				pub enum CompatV3ToV4 {

				    V3(v3::V3Reader),

				    Compat(CompatV2ToV3),

				}

				impl CompatV3ToV4 {

				    pub fn new(v3: v3::V3Reader) -> CompatV3ToV4 {

				        CompatV3ToV4::V3(v3)

				    }

				    pub fn to_v5(self) -> CompatV4ToV5 {

				        CompatV4ToV5::Compat(self)

				    }

				    pub fn version(&self) -> crate::Version {

				        match self {

				            CompatV3ToV4::V3(v3) => v3.version(),

				            CompatV3ToV4::Compat(compat) => compat.version(),

				        }

				    }

				    pub fn date(&self) -> Option<time::OffsetDateTime> {

				        match self {

				            CompatV3ToV4::V3(v3) => v3.date(),

				            CompatV3ToV4::Compat(compat) => compat.date(),

				        }

				    }

				    pub fn instance_uid(&self) -> Result<Option<uuid::Uuid>> {

				        Ok(None)

				    }

				    pub fn indexes(&self) -> Result<impl Iterator<Item = Result<CompatIndexV3ToV4>> + '_> {

				        Ok(match self {

				            CompatV3ToV4::V3(v3) => {

				                Box::new(v3.indexes()?.map(|index| index.map(CompatIndexV3ToV4::from)))

				                    as Box<dyn Iterator<Item = Result<CompatIndexV3ToV4>> + '_>

				            }

				            CompatV3ToV4::Compat(compat) => {

				                Box::new(compat.indexes()?.map(|index| index.map(CompatIndexV3ToV4::from)))

				                    as Box<dyn Iterator<Item = Result<CompatIndexV3ToV4>> + '_>

				            }

				        })

				    }

				    pub fn tasks(

				        &mut self,

				    ) -> Box<dyn Iterator<Item = Result<(v4::Task, Option<Box<UpdateFile>>)>> + '_> {

				        let indexes = match self {

				            CompatV3ToV4::V3(v3) => v3.index_uuid(),

				            CompatV3ToV4::Compat(compat) => compat.index_uuid(),

				        };

				        let tasks = match self {

				            CompatV3ToV4::V3(v3) => v3.tasks(),

				            CompatV3ToV4::Compat(compat) => compat.tasks(),

				        };

				        Box::new(

				            tasks

				                // we need to override the old task ids that were generated

				                // by index in favor of a global unique incremental ID.

				                .enumerate()

				                .map(move |(task_id, task)| {

				                    task.map(|(task, content_file)| {

				                        let index_uid = indexes

				                            .iter()

				                            .find(|index| index.uuid == task.uuid)

				                            .map(|index| index.uid.clone());

				                        let index_uid = match index_uid {

				                            Some(uid) => uid,

				                            None => {

				                                tracing::warn!(

				                                    "Error while importing the update {}.",

				                                    task.update.id()

				                                );

				                                tracing::warn!(

				                                    "The index associated to the uuid `{}` could not be retrieved.",

				                                    task.uuid.to_string()

				                                );

				                                if task.update.is_finished() {

				                                    // we're fucking with his history but not his data, that's ok-ish.

				                                    tracing::warn!("The index-uuid will be set as `unknown`.");

				                                    String::from("unknown")

				                                } else {

				                                    tracing::warn!("The task will be ignored.");

				                                    return None;

				                                }

				                            }

				                        };

				                        let task = v4::Task {

				                            id: task_id as u32,

				                            index_uid: v4::meta::IndexUid(index_uid),

				                            content: match task.update.meta() {

				                                v3::Kind::DeleteDocuments(documents) => {

				                                    v4::tasks::TaskContent::DocumentDeletion(

				                                        v4::tasks::DocumentDeletion::Ids(documents.clone()),

				                                    )

				                                }

				                                v3::Kind::DocumentAddition {

				                                    primary_key,

				                                    method,

				                                    content_uuid,

				                                } => v4::tasks::TaskContent::DocumentAddition {

				                                    merge_strategy: match method {

				                                        v3::updates::IndexDocumentsMethod::ReplaceDocuments => {

				                                            v4::tasks::IndexDocumentsMethod::ReplaceDocuments

				                                        }

				                                        v3::updates::IndexDocumentsMethod::UpdateDocuments => {

				                                            v4::tasks::IndexDocumentsMethod::UpdateDocuments

				                                        }

				                                    },

				                                    primary_key: primary_key.clone(),

				                                    documents_count: 0, // we don't have this info

				                                    allow_index_creation: true, // there was no API-key in the v3

				                                    content_uuid: *content_uuid,

				                                },

				                                v3::Kind::Settings(settings) => {

				                                    v4::tasks::TaskContent::SettingsUpdate {

				                                        settings: v4::Settings::from(settings.clone()),

				                                        is_deletion: false, // that didn't exist at this time

				                                        allow_index_creation: true, // there was no API-key in the v3

				                                    }

				                                }

				                                v3::Kind::ClearDocuments => {

				                                    v4::tasks::TaskContent::DocumentDeletion(

				                                        v4::tasks::DocumentDeletion::Clear,

				                                    )

				                                }

				                            },

				                            events: match task.update {

				                                v3::Status::Processing(processing) => {

				                                    vec![v4::tasks::TaskEvent::Created(processing.from.enqueued_at)]

				                                }

				                                v3::Status::Enqueued(enqueued) => {

				                                    vec![v4::tasks::TaskEvent::Created(enqueued.enqueued_at)]

				                                }

				                                v3::Status::Processed(processed) => {

				                                    vec![

				                                        v4::tasks::TaskEvent::Created(

				                                            processed.from.from.enqueued_at,

				                                        ),

				                                        v4::tasks::TaskEvent::Processing(

				                                            processed.from.started_processing_at,

				                                        ),

				                                        v4::tasks::TaskEvent::Succeded {

				                                            result: match processed.success {

				                                                v3::updates::UpdateResult::DocumentsAddition(

				                                                    document_addition,

				                                                ) => v4::tasks::TaskResult::DocumentAddition {

				                                                    indexed_documents: document_addition

				                                                        .nb_documents

				                                                        as u64,

				                                                },

				                                                v3::updates::UpdateResult::DocumentDeletion {

				                                                    deleted,

				                                                } => v4::tasks::TaskResult::DocumentDeletion {

				                                                    deleted_documents: deleted,

				                                                },

				                                                v3::updates::UpdateResult::Other => {

				                                                    v4::tasks::TaskResult::Other

				                                                }

				                                            },

				                                            timestamp: processed.processed_at,

				                                        },

				                                    ]

				                                }

				                                v3::Status::Failed(failed) => vec![

				                                    v4::tasks::TaskEvent::Created(failed.from.from.enqueued_at),

				                                    v4::tasks::TaskEvent::Processing(

				                                        failed.from.started_processing_at,

				                                    ),

				                                    v4::tasks::TaskEvent::Failed {

				                                        error: v4::ResponseError::from_msg(

				                                            failed.msg.to_string(),

				                                            failed.code.into(),

				                                        ),

				                                        timestamp: failed.failed_at,

				                                    },

				                                ],

				                                v3::Status::Aborted(aborted) => vec![

				                            v4::tasks::TaskEvent::Created(aborted.from.enqueued_at),

				                            v4::tasks::TaskEvent::Failed {

				                                error: v4::ResponseError::from_msg(

				                                    "Task was aborted in a previous version of meilisearch."

				                                        .to_string(),

				                                    v4::errors::Code::UnretrievableErrorCode,

				                                ),

				                                timestamp: aborted.aborted_at,

				                            },

				                        ],

				                            },

				                        };

				                        Some((task, content_file))

				                    })

				                })

				                .filter_map(|res| res.transpose()),

				        )

				    }

				    pub fn keys(&mut self) -> Box<dyn Iterator<Item = Result<v4::Key>> + '_> {

				        Box::new(std::iter::empty())

				    }

				}

				pub enum CompatIndexV3ToV4 {

				    V3(v3::V3IndexReader),

				    Compat(CompatIndexV2ToV3),

				}

				impl From<v3::V3IndexReader> for CompatIndexV3ToV4 {

				    fn from(index_reader: v3::V3IndexReader) -> Self {

				        Self::V3(index_reader)

				    }

				}

				impl From<CompatIndexV2ToV3> for CompatIndexV3ToV4 {

				    fn from(index_reader: CompatIndexV2ToV3) -> Self {

				        Self::Compat(index_reader)

				    }

				}

				impl CompatIndexV3ToV4 {

				    pub fn new(v3: v3::V3IndexReader) -> CompatIndexV3ToV4 {

				        CompatIndexV3ToV4::V3(v3)

				    }

				    pub fn metadata(&self) -> &crate::IndexMetadata {

				        match self {

				            CompatIndexV3ToV4::V3(v3) => v3.metadata(),

				            CompatIndexV3ToV4::Compat(compat) => compat.metadata(),

				        }

				    }

				    pub fn documents(&mut self) -> Result<Box<dyn Iterator<Item = Result<v4::Document>> + '_>> {

				        match self {

				            CompatIndexV3ToV4::V3(v3) => v3

				                .documents()

				                .map(|iter| Box::new(iter) as Box<dyn Iterator<Item = Result<v4::Document>> + '_>),

				            CompatIndexV3ToV4::Compat(compat) => compat

				                .documents()

				                .map(|iter| Box::new(iter) as Box<dyn Iterator<Item = Result<v4::Document>> + '_>),

				        }

				    }

				    pub fn settings(&mut self) -> Result<v4::Settings<v4::Checked>> {

				        Ok(match self {

				            CompatIndexV3ToV4::V3(v3) => {

				                v4::Settings::<v4::Unchecked>::from(v3.settings()?).check()

				            }

				            CompatIndexV3ToV4::Compat(compat) => {

				                v4::Settings::<v4::Unchecked>::from(compat.settings()?).check()

				            }

				        })

				    }

				}

				impl<T> From<v3::Setting<T>> for v4::Setting<T> {

				    fn from(setting: v3::Setting<T>) -> Self {

				        match setting {

				            v3::Setting::Set(t) => v4::Setting::Set(t),

				            v3::Setting::Reset => v4::Setting::Reset,

				            v3::Setting::NotSet => v4::Setting::NotSet,

				        }

				    }

				}

				impl From<v3::Code> for v4::Code {

				    fn from(code: v3::Code) -> Self {

				        match code {

				            v3::Code::CreateIndex => v4::Code::CreateIndex,

				            v3::Code::IndexAlreadyExists => v4::Code::IndexAlreadyExists,

				            v3::Code::IndexNotFound => v4::Code::IndexNotFound,

				            v3::Code::InvalidIndexUid => v4::Code::InvalidIndexUid,

				            v3::Code::InvalidState => v4::Code::InvalidState,

				            v3::Code::MissingPrimaryKey => v4::Code::MissingPrimaryKey,

				            v3::Code::PrimaryKeyAlreadyPresent => v4::Code::PrimaryKeyAlreadyPresent,

				            v3::Code::MaxFieldsLimitExceeded => v4::Code::MaxFieldsLimitExceeded,

				            v3::Code::MissingDocumentId => v4::Code::MissingDocumentId,

				            v3::Code::InvalidDocumentId => v4::Code::InvalidDocumentId,

				            v3::Code::Filter => v4::Code::Filter,

				            v3::Code::Sort => v4::Code::Sort,

				            v3::Code::BadParameter => v4::Code::BadParameter,

				            v3::Code::BadRequest => v4::Code::BadRequest,

				            v3::Code::DatabaseSizeLimitReached => v4::Code::DatabaseSizeLimitReached,

				            v3::Code::DocumentNotFound => v4::Code::DocumentNotFound,

				            v3::Code::Internal => v4::Code::Internal,

				            v3::Code::InvalidGeoField => v4::Code::InvalidGeoField,

				            v3::Code::InvalidRankingRule => v4::Code::InvalidRankingRule,

				            v3::Code::InvalidStore => v4::Code::InvalidStore,

				            v3::Code::InvalidToken => v4::Code::InvalidToken,

				            v3::Code::MissingAuthorizationHeader => v4::Code::MissingAuthorizationHeader,

				            v3::Code::NoSpaceLeftOnDevice => v4::Code::NoSpaceLeftOnDevice,

				            v3::Code::DumpNotFound => v4::Code::DumpNotFound,

				            v3::Code::TaskNotFound => v4::Code::TaskNotFound,

				            v3::Code::PayloadTooLarge => v4::Code::PayloadTooLarge,

				            v3::Code::RetrieveDocument => v4::Code::RetrieveDocument,

				            v3::Code::SearchDocuments => v4::Code::SearchDocuments,

				            v3::Code::UnsupportedMediaType => v4::Code::UnsupportedMediaType,

				            v3::Code::DumpAlreadyInProgress => v4::Code::DumpAlreadyInProgress,

				            v3::Code::DumpProcessFailed => v4::Code::DumpProcessFailed,

				            v3::Code::InvalidContentType => v4::Code::InvalidContentType,

				            v3::Code::MissingContentType => v4::Code::MissingContentType,

				            v3::Code::MalformedPayload => v4::Code::MalformedPayload,

				            v3::Code::MissingPayload => v4::Code::MissingPayload,

				            v3::Code::UnretrievableErrorCode => v4::Code::UnretrievableErrorCode,

				            v3::Code::MalformedDump => v4::Code::MalformedDump,

				        }

				    }

				}

				impl<T> From<v3::Settings<T>> for v4::Settings<v4::Unchecked> {

				    fn from(settings: v3::Settings<T>) -> Self {

				        v4::Settings {

				            displayed_attributes: settings.displayed_attributes.into(),

				            searchable_attributes: settings.searchable_attributes.into(),

				            filterable_attributes: settings.filterable_attributes.into(),

				            sortable_attributes: settings.sortable_attributes.into(),

				            ranking_rules: settings.ranking_rules.into(),

				            stop_words: settings.stop_words.into(),

				            synonyms: settings.synonyms.into(),

				            distinct_attribute: settings.distinct_attribute.into(),

				            typo_tolerance: v4::Setting::NotSet,

				            _kind: std::marker::PhantomData,

				        }

				    }

				}

				#[cfg(test)]

				pub(crate) mod test {

				    use std::fs::File;

				    use std::io::BufReader;

				    use flate2::bufread::GzDecoder;

				    use meili_snap::insta;

				    use tempfile::TempDir;

				    use super::*;

				    #[test]

				    fn compat_v3_v4() {

				        let dump = File::open("tests/assets/v3.dump").unwrap();

				        let dir = TempDir::new().unwrap();

				        let mut dump = BufReader::new(dump);

				        let gz = GzDecoder::new(&mut dump);

				        let mut archive = tar::Archive::new(gz);

				        archive.unpack(dir.path()).unwrap();

				        let mut dump = v3::V3Reader::open(dir).unwrap().to_v4();

				        // top level infos

				        insta::assert_display_snapshot!(dump.date().unwrap(), @"2022-10-07 11:39:03.709153554 +00:00:00");

				        // tasks

				        let tasks = dump.tasks().collect::<Result<Vec<_>>>().unwrap();

				        let (tasks, mut update_files): (Vec<_>, Vec<_>) = tasks.into_iter().unzip();

				        meili_snap::snapshot_hash!(meili_snap::json_string!(tasks), @"79bc053583a1a7172bbaaafb1edaeb78");

				        assert_eq!(update_files.len(), 10);

				        assert!(update_files[0].is_some()); // the enqueued document addition

				        assert!(update_files[1..].iter().all(|u| u.is_none())); // everything already processed

				        let update_file = update_files.remove(0).unwrap().collect::<Result<Vec<_>>>().unwrap();

				        meili_snap::snapshot_hash!(meili_snap::json_string!(update_file), @"7b8889539b669c7b9ddba448bafa385d");

				        // keys

				        let keys = dump.keys().collect::<Result<Vec<_>>>().unwrap();

				        meili_snap::snapshot_hash!(meili_snap::json_string!(keys, { "[].uid" => "[uuid]" }), @"d751713988987e9331980363e24189ce");

				        // indexes

				        let mut indexes = dump.indexes().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        // the index are not ordered in any way by default

				        indexes.sort_by_key(|index| index.metadata().uid.to_string());

				        let mut products = indexes.pop().unwrap();

				        let mut movies2 = indexes.pop().unwrap();

				        let mut movies = indexes.pop().unwrap();

				        let mut spells = indexes.pop().unwrap();

				        assert!(indexes.is_empty());

				        // products

				        insta::assert_json_snapshot!(products.metadata(), { ".createdAt" => "[now]", ".updatedAt" => "[now]" }, @r###"

				        {

				          "uid": "products",

				          "primaryKey": "sku",

				          "createdAt": "[now]",

				          "updatedAt": "[now]"

				        }

				        "###);

				        insta::assert_json_snapshot!(products.settings().unwrap());

				        let documents = products.documents().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        assert_eq!(documents.len(), 10);

				        meili_snap::snapshot_hash!(format!("{:#?}", documents), @"548284a84de510f71e88e6cdea495cf5");

				        // movies

				        insta::assert_json_snapshot!(movies.metadata(), { ".createdAt" => "[now]", ".updatedAt" => "[now]" }, @r###"

				        {

				          "uid": "movies",

				          "primaryKey": "id",

				          "createdAt": "[now]",

				          "updatedAt": "[now]"

				        }

				        "###);

				        insta::assert_json_snapshot!(movies.settings().unwrap());

				        let documents = movies.documents().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        assert_eq!(documents.len(), 110);

				        meili_snap::snapshot_hash!(format!("{:#?}", documents), @"d153b5a81d8b3cdcbe1dec270b574022");

				        // movies2

				        insta::assert_json_snapshot!(movies2.metadata(), { ".createdAt" => "[now]", ".updatedAt" => "[now]" }, @r###"

				        {

				          "uid": "movies_2",

				          "primaryKey": null,

				          "createdAt": "[now]",

				          "updatedAt": "[now]"

				        }

				        "###);

				        insta::assert_json_snapshot!(movies2.settings().unwrap());

				        let documents = movies2.documents().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        assert_eq!(documents.len(), 0);

				        meili_snap::snapshot_hash!(format!("{:#?}", documents), @"d751713988987e9331980363e24189ce");

				        // spells

				        insta::assert_json_snapshot!(spells.metadata(), { ".createdAt" => "[now]", ".updatedAt" => "[now]" }, @r###"

				        {

				          "uid": "dnd_spells",

				          "primaryKey": "index",

				          "createdAt": "[now]",

				          "updatedAt": "[now]"

				        }

				        "###);

				        insta::assert_json_snapshot!(spells.settings().unwrap());

				        let documents = spells.documents().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        assert_eq!(documents.len(), 10);

				        meili_snap::snapshot_hash!(format!("{:#?}", documents), @"235016433dd04262c7f2da01d1e808ce");

				    }

				}

									
										467

dump/src/reader/compat/v4_to_v5.rs
									
										Normal file
									
												View File
												
				@@ -0,0 +1,467 @@

				use super::v3_to_v4::{CompatIndexV3ToV4, CompatV3ToV4};

				use super::v5_to_v6::CompatV5ToV6;

				use crate::reader::{v4, v5, Document};

				use crate::Result;

				pub enum CompatV4ToV5 {

				    V4(v4::V4Reader),

				    Compat(CompatV3ToV4),

				}

				impl CompatV4ToV5 {

				    pub fn new(v4: v4::V4Reader) -> CompatV4ToV5 {

				        CompatV4ToV5::V4(v4)

				    }

				    pub fn to_v6(self) -> CompatV5ToV6 {

				        CompatV5ToV6::Compat(self)

				    }

				    pub fn version(&self) -> crate::Version {

				        match self {

				            CompatV4ToV5::V4(v4) => v4.version(),

				            CompatV4ToV5::Compat(compat) => compat.version(),

				        }

				    }

				    pub fn date(&self) -> Option<time::OffsetDateTime> {

				        match self {

				            CompatV4ToV5::V4(v4) => v4.date(),

				            CompatV4ToV5::Compat(compat) => compat.date(),

				        }

				    }

				    pub fn instance_uid(&self) -> Result<Option<uuid::Uuid>> {

				        match self {

				            CompatV4ToV5::V4(v4) => v4.instance_uid(),

				            CompatV4ToV5::Compat(compat) => compat.instance_uid(),

				        }

				    }

				    pub fn indexes(&self) -> Result<Box<dyn Iterator<Item = Result<CompatIndexV4ToV5>> + '_>> {

				        Ok(match self {

				            CompatV4ToV5::V4(v4) => {

				                Box::new(v4.indexes()?.map(|index| index.map(CompatIndexV4ToV5::from)))

				                    as Box<dyn Iterator<Item = Result<CompatIndexV4ToV5>> + '_>

				            }

				            CompatV4ToV5::Compat(compat) => {

				                Box::new(compat.indexes()?.map(|index| index.map(CompatIndexV4ToV5::from)))

				                    as Box<dyn Iterator<Item = Result<CompatIndexV4ToV5>> + '_>

				            }

				        })

				    }

				    pub fn tasks(

				        &mut self,

				    ) -> Box<dyn Iterator<Item = Result<(v5::Task, Option<Box<crate::reader::UpdateFile>>)>> + '_>

				    {

				        let tasks = match self {

				            CompatV4ToV5::V4(v4) => v4.tasks(),

				            CompatV4ToV5::Compat(compat) => compat.tasks(),

				        };

				        Box::new(tasks.map(|task| {

				            task.map(|(task, content_file)| {

				                let task = v5::Task {

				                    id: task.id,

				                    content: match task.content {

				                        v4::tasks::TaskContent::DocumentAddition {

				                            content_uuid,

				                            merge_strategy,

				                            primary_key,

				                            documents_count,

				                            allow_index_creation,

				                        } => v5::tasks::TaskContent::DocumentAddition {

				                            index_uid: v5::meta::IndexUid(task.index_uid.0),

				                            content_uuid,

				                            merge_strategy: match merge_strategy {

				                                v4::tasks::IndexDocumentsMethod::ReplaceDocuments => {

				                                    v5::tasks::IndexDocumentsMethod::ReplaceDocuments

				                                }

				                                v4::tasks::IndexDocumentsMethod::UpdateDocuments => {

				                                    v5::tasks::IndexDocumentsMethod::UpdateDocuments

				                                }

				                            },

				                            primary_key,

				                            documents_count,

				                            allow_index_creation,

				                        },

				                        v4::tasks::TaskContent::DocumentDeletion(deletion) => {

				                            v5::tasks::TaskContent::DocumentDeletion {

				                                index_uid: v5::meta::IndexUid(task.index_uid.0),

				                                deletion: match deletion {

				                                    v4::tasks::DocumentDeletion::Clear => {

				                                        v5::tasks::DocumentDeletion::Clear

				                                    }

				                                    v4::tasks::DocumentDeletion::Ids(ids) => {

				                                        v5::tasks::DocumentDeletion::Ids(ids)

				                                    }

				                                },

				                            }

				                        }

				                        v4::tasks::TaskContent::SettingsUpdate {

				                            settings,

				                            is_deletion,

				                            allow_index_creation,

				                        } => v5::tasks::TaskContent::SettingsUpdate {

				                            index_uid: v5::meta::IndexUid(task.index_uid.0),

				                            settings: settings.into(),

				                            is_deletion,

				                            allow_index_creation,

				                        },

				                        v4::tasks::TaskContent::IndexDeletion => {

				                            v5::tasks::TaskContent::IndexDeletion {

				                                index_uid: v5::meta::IndexUid(task.index_uid.0),

				                            }

				                        }

				                        v4::tasks::TaskContent::IndexCreation { primary_key } => {

				                            v5::tasks::TaskContent::IndexCreation {

				                                index_uid: v5::meta::IndexUid(task.index_uid.0),

				                                primary_key,

				                            }

				                        }

				                        v4::tasks::TaskContent::IndexUpdate { primary_key } => {

				                            v5::tasks::TaskContent::IndexUpdate {

				                                index_uid: v5::meta::IndexUid(task.index_uid.0),

				                                primary_key,

				                            }

				                        }

				                    },

				                    events: task

				                        .events

				                        .into_iter()

				                        .map(|event| match event {

				                            v4::tasks::TaskEvent::Created(date) => {

				                                v5::tasks::TaskEvent::Created(date)

				                            }

				                            v4::tasks::TaskEvent::Batched { timestamp, batch_id } => {

				                                v5::tasks::TaskEvent::Batched { timestamp, batch_id }

				                            }

				                            v4::tasks::TaskEvent::Processing(date) => {

				                                v5::tasks::TaskEvent::Processing(date)

				                            }

				                            v4::tasks::TaskEvent::Succeded { result, timestamp } => {

				                                v5::tasks::TaskEvent::Succeeded {

				                                    result: match result {

				                                        v4::tasks::TaskResult::DocumentAddition {

				                                            indexed_documents,

				                                        } => v5::tasks::TaskResult::DocumentAddition {

				                                            indexed_documents,

				                                        },

				                                        v4::tasks::TaskResult::DocumentDeletion {

				                                            deleted_documents,

				                                        } => v5::tasks::TaskResult::DocumentDeletion {

				                                            deleted_documents,

				                                        },

				                                        v4::tasks::TaskResult::ClearAll { deleted_documents } => {

				                                            v5::tasks::TaskResult::ClearAll { deleted_documents }

				                                        }

				                                        v4::tasks::TaskResult::Other => {

				                                            v5::tasks::TaskResult::Other

				                                        }

				                                    },

				                                    timestamp,

				                                }

				                            }

				                            v4::tasks::TaskEvent::Failed { error, timestamp } => {

				                                v5::tasks::TaskEvent::Failed {

				                                    error: v5::ResponseError::from(error),

				                                    timestamp,

				                                }

				                            }

				                        })

				                        .collect(),

				                };

				                (task, content_file)

				            })

				        }))

				    }

				    pub fn keys(&mut self) -> Box<dyn Iterator<Item = Result<v5::Key>> + '_> {

				        let keys = match self {

				            CompatV4ToV5::V4(v4) => v4.keys(),

				            CompatV4ToV5::Compat(compat) => compat.keys(),

				        };

				        Box::new(keys.map(|key| {

				            key.map(|key| v5::Key {

				                description: key.description,

				                name: None,

				                uid: v5::keys::KeyId::new_v4(),

				                actions: key.actions.into_iter().filter_map(|action| action.into()).collect(),

				                indexes: key

				                    .indexes

				                    .into_iter()

				                    .map(|index| match index.as_str() {

				                        "*" => v5::StarOr::Star,

				                        _ => v5::StarOr::Other(v5::meta::IndexUid(index)),

				                    })

				                    .collect(),

				                expires_at: key.expires_at,

				                created_at: key.created_at,

				                updated_at: key.updated_at,

				            })

				        }))

				    }

				}

				pub enum CompatIndexV4ToV5 {

				    V4(v4::V4IndexReader),

				    Compat(CompatIndexV3ToV4),

				}

				impl From<v4::V4IndexReader> for CompatIndexV4ToV5 {

				    fn from(index_reader: v4::V4IndexReader) -> Self {

				        Self::V4(index_reader)

				    }

				}

				impl From<CompatIndexV3ToV4> for CompatIndexV4ToV5 {

				    fn from(index_reader: CompatIndexV3ToV4) -> Self {

				        Self::Compat(index_reader)

				    }

				}

				impl CompatIndexV4ToV5 {

				    pub fn metadata(&self) -> &crate::IndexMetadata {

				        match self {

				            CompatIndexV4ToV5::V4(v4) => v4.metadata(),

				            CompatIndexV4ToV5::Compat(compat) => compat.metadata(),

				        }

				    }

				    pub fn documents(&mut self) -> Result<Box<dyn Iterator<Item = Result<Document>> + '_>> {

				        match self {

				            CompatIndexV4ToV5::V4(v4) => v4

				                .documents()

				                .map(|iter| Box::new(iter) as Box<dyn Iterator<Item = Result<Document>> + '_>),

				            CompatIndexV4ToV5::Compat(compat) => compat

				                .documents()

				                .map(|iter| Box::new(iter) as Box<dyn Iterator<Item = Result<Document>> + '_>),

				        }

				    }

				    pub fn settings(&mut self) -> Result<v5::Settings<v5::Checked>> {

				        match self {

				            CompatIndexV4ToV5::V4(v4) => Ok(v5::Settings::from(v4.settings()?).check()),

				            CompatIndexV4ToV5::Compat(compat) => Ok(v5::Settings::from(compat.settings()?).check()),

				        }

				    }

				}

				impl<T> From<v4::Setting<T>> for v5::Setting<T> {

				    fn from(setting: v4::Setting<T>) -> Self {

				        match setting {

				            v4::Setting::Set(t) => v5::Setting::Set(t),

				            v4::Setting::Reset => v5::Setting::Reset,

				            v4::Setting::NotSet => v5::Setting::NotSet,

				        }

				    }

				}

				impl From<v4::ResponseError> for v5::ResponseError {

				    fn from(error: v4::ResponseError) -> Self {

				        let code = match error.error_code.as_ref() {

				            "index_creation_failed" => v5::Code::CreateIndex,

				            "index_already_exists" => v5::Code::IndexAlreadyExists,

				            "index_not_found" => v5::Code::IndexNotFound,

				            "invalid_index_uid" => v5::Code::InvalidIndexUid,

				            "invalid_min_word_length_for_typo" => v5::Code::InvalidMinWordLengthForTypo,

				            "invalid_state" => v5::Code::InvalidState,

				            "primary_key_inference_failed" => v5::Code::MissingPrimaryKey,

				            "index_primary_key_already_exists" => v5::Code::PrimaryKeyAlreadyPresent,

				            "max_fields_limit_exceeded" => v5::Code::MaxFieldsLimitExceeded,

				            "missing_document_id" => v5::Code::MissingDocumentId,

				            "invalid_document_id" => v5::Code::InvalidDocumentId,

				            "invalid_filter" => v5::Code::Filter,

				            "invalid_sort" => v5::Code::Sort,

				            "bad_parameter" => v5::Code::BadParameter,

				            "bad_request" => v5::Code::BadRequest,

				            "database_size_limit_reached" => v5::Code::DatabaseSizeLimitReached,

				            "document_not_found" => v5::Code::DocumentNotFound,

				            "internal" => v5::Code::Internal,

				            "invalid_geo_field" => v5::Code::InvalidGeoField,

				            "invalid_ranking_rule" => v5::Code::InvalidRankingRule,

				            "invalid_store_file" => v5::Code::InvalidStore,

				            "invalid_api_key" => v5::Code::InvalidToken,

				            "missing_authorization_header" => v5::Code::MissingAuthorizationHeader,

				            "no_space_left_on_device" => v5::Code::NoSpaceLeftOnDevice,

				            "dump_not_found" => v5::Code::DumpNotFound,

				            "task_not_found" => v5::Code::TaskNotFound,

				            "payload_too_large" => v5::Code::PayloadTooLarge,

				            "unretrievable_document" => v5::Code::RetrieveDocument,

				            "search_error" => v5::Code::SearchDocuments,

				            "unsupported_media_type" => v5::Code::UnsupportedMediaType,

				            "dump_already_processing" => v5::Code::DumpAlreadyInProgress,

				            "dump_process_failed" => v5::Code::DumpProcessFailed,

				            "invalid_content_type" => v5::Code::InvalidContentType,

				            "missing_content_type" => v5::Code::MissingContentType,

				            "malformed_payload" => v5::Code::MalformedPayload,

				            "missing_payload" => v5::Code::MissingPayload,

				            "api_key_not_found" => v5::Code::ApiKeyNotFound,

				            "missing_parameter" => v5::Code::MissingParameter,

				            "invalid_api_key_actions" => v5::Code::InvalidApiKeyActions,

				            "invalid_api_key_indexes" => v5::Code::InvalidApiKeyIndexes,

				            "invalid_api_key_expires_at" => v5::Code::InvalidApiKeyExpiresAt,

				            "invalid_api_key_description" => v5::Code::InvalidApiKeyDescription,

				            other => {

				                tracing::warn!("Unknown error code {}", other);

				                v5::Code::UnretrievableErrorCode

				            }

				        };

				        v5::ResponseError::from_msg(error.message, code)

				    }

				}

				impl<T> From<v4::Settings<T>> for v5::Settings<v5::Unchecked> {

				    fn from(settings: v4::Settings<T>) -> Self {

				        v5::Settings {

				            displayed_attributes: settings.displayed_attributes.into(),

				            searchable_attributes: settings.searchable_attributes.into(),

				            filterable_attributes: settings.filterable_attributes.into(),

				            sortable_attributes: settings.sortable_attributes.into(),

				            ranking_rules: settings.ranking_rules.into(),

				            stop_words: settings.stop_words.into(),

				            synonyms: settings.synonyms.into(),

				            distinct_attribute: settings.distinct_attribute.into(),

				            typo_tolerance: match settings.typo_tolerance {

				                v4::Setting::Set(typo) => v5::Setting::Set(v5::TypoTolerance {

				                    enabled: typo.enabled.into(),

				                    min_word_size_for_typos: match typo.min_word_size_for_typos {

				                        v4::Setting::Set(t) => v5::Setting::Set(v5::MinWordSizeForTypos {

				                            one_typo: t.one_typo.into(),

				                            two_typos: t.two_typos.into(),

				                        }),

				                        v4::Setting::Reset => v5::Setting::Reset,

				                        v4::Setting::NotSet => v5::Setting::NotSet,

				                    },

				                    disable_on_words: typo.disable_on_words.into(),

				                    disable_on_attributes: typo.disable_on_attributes.into(),

				                }),

				                v4::Setting::Reset => v5::Setting::Reset,

				                v4::Setting::NotSet => v5::Setting::NotSet,

				            },

				            faceting: v5::Setting::NotSet,

				            pagination: v5::Setting::NotSet,

				            _kind: std::marker::PhantomData,

				        }

				    }

				}

				impl From<v4::Action> for Option<v5::Action> {

				    fn from(key: v4::Action) -> Self {

				        match key {

				            v4::Action::All => Some(v5::Action::All),

				            v4::Action::Search => Some(v5::Action::Search),

				            v4::Action::DocumentsAdd => Some(v5::Action::DocumentsAdd),

				            v4::Action::DocumentsGet => Some(v5::Action::DocumentsGet),

				            v4::Action::DocumentsDelete => Some(v5::Action::DocumentsDelete),

				            v4::Action::IndexesAdd => Some(v5::Action::IndexesAdd),

				            v4::Action::IndexesGet => Some(v5::Action::IndexesGet),

				            v4::Action::IndexesUpdate => Some(v5::Action::IndexesUpdate),

				            v4::Action::IndexesDelete => Some(v5::Action::IndexesDelete),

				            v4::Action::TasksGet => Some(v5::Action::TasksGet),

				            v4::Action::SettingsGet => Some(v5::Action::SettingsGet),

				            v4::Action::SettingsUpdate => Some(v5::Action::SettingsUpdate),

				            v4::Action::StatsGet => Some(v5::Action::StatsGet),

				            v4::Action::DumpsCreate => Some(v5::Action::DumpsCreate),

				            v4::Action::DumpsGet => None,

				            v4::Action::Version => Some(v5::Action::Version),

				        }

				    }

				}

				#[cfg(test)]

				pub(crate) mod test {

				    use std::fs::File;

				    use std::io::BufReader;

				    use flate2::bufread::GzDecoder;

				    use meili_snap::insta;

				    use tempfile::TempDir;

				    use super::*;

				    #[test]

				    fn compat_v4_v5() {

				        let dump = File::open("tests/assets/v4.dump").unwrap();

				        let dir = TempDir::new().unwrap();

				        let mut dump = BufReader::new(dump);

				        let gz = GzDecoder::new(&mut dump);

				        let mut archive = tar::Archive::new(gz);

				        archive.unpack(dir.path()).unwrap();

				        let mut dump = v4::V4Reader::open(dir).unwrap().to_v5();

				        // top level infos

				        insta::assert_display_snapshot!(dump.date().unwrap(), @"2022-10-06 12:53:49.131989609 +00:00:00");

				        insta::assert_display_snapshot!(dump.instance_uid().unwrap().unwrap(), @"9e15e977-f2ae-4761-943f-1eaf75fd736d");

				        // tasks

				        let tasks = dump.tasks().collect::<Result<Vec<_>>>().unwrap();

				        let (tasks, update_files): (Vec<_>, Vec<_>) = tasks.into_iter().unzip();

				        meili_snap::snapshot_hash!(meili_snap::json_string!(tasks), @"ed9a30cded4c046ef46f7cff7450347e");

				        assert_eq!(update_files.len(), 10);

				        assert!(update_files[0].is_some()); // the enqueued document addition

				        assert!(update_files[1..].iter().all(|u| u.is_none())); // everything already processed

				        // keys

				        let keys = dump.keys().collect::<Result<Vec<_>>>().unwrap();

				        meili_snap::snapshot_hash!(meili_snap::json_string!(keys, { "[].uid" => "[uuid]" }), @"1384361d734fd77c23804c9696228660");

				        // indexes

				        let mut indexes = dump.indexes().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        // the index are not ordered in any way by default

				        indexes.sort_by_key(|index| index.metadata().uid.to_string());

				        let mut products = indexes.pop().unwrap();

				        let mut movies = indexes.pop().unwrap();

				        let mut spells = indexes.pop().unwrap();

				        assert!(indexes.is_empty());

				        // products

				        insta::assert_json_snapshot!(products.metadata(), { ".createdAt" => "[now]", ".updatedAt" => "[now]" }, @r###"

				        {

				          "uid": "products",

				          "primaryKey": "sku",

				          "createdAt": "[now]",

				          "updatedAt": "[now]"

				        }

				        "###);

				        insta::assert_json_snapshot!(products.settings().unwrap());

				        let documents = products.documents().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        assert_eq!(documents.len(), 10);

				        meili_snap::snapshot_hash!(format!("{:#?}", documents), @"b01c8371aea4c7171af0d4d846a2bdca");

				        // movies

				        insta::assert_json_snapshot!(movies.metadata(), { ".createdAt" => "[now]", ".updatedAt" => "[now]" }, @r###"

				        {

				          "uid": "movies",

				          "primaryKey": "id",

				          "createdAt": "[now]",

				          "updatedAt": "[now]"

				        }

				        "###);

				        insta::assert_json_snapshot!(movies.settings().unwrap());

				        let documents = movies.documents().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        assert_eq!(documents.len(), 110);

				        meili_snap::snapshot_hash!(format!("{:#?}", documents), @"786022a66ecb992c8a2a60fee070a5ab");

				        // spells

				        insta::assert_json_snapshot!(spells.metadata(), { ".createdAt" => "[now]", ".updatedAt" => "[now]" }, @r###"

				        {

				          "uid": "dnd_spells",

				          "primaryKey": "index",

				          "createdAt": "[now]",

				          "updatedAt": "[now]"

				        }

				        "###);

				        insta::assert_json_snapshot!(spells.settings().unwrap());

				        let documents = spells.documents().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        assert_eq!(documents.len(), 10);

				        meili_snap::snapshot_hash!(format!("{:#?}", documents), @"235016433dd04262c7f2da01d1e808ce");

				    }

				}

									
										515

dump/src/reader/compat/v5_to_v6.rs
									
										Normal file
									
												View File
												
				@@ -0,0 +1,515 @@

				use std::str::FromStr;

				use super::v4_to_v5::{CompatIndexV4ToV5, CompatV4ToV5};

				use crate::reader::{v5, v6, Document, UpdateFile};

				use crate::Result;

				pub enum CompatV5ToV6 {

				    V5(v5::V5Reader),

				    Compat(CompatV4ToV5),

				}

				impl CompatV5ToV6 {

				    pub fn new_v5(v5: v5::V5Reader) -> CompatV5ToV6 {

				        CompatV5ToV6::V5(v5)

				    }

				    pub fn version(&self) -> crate::Version {

				        match self {

				            CompatV5ToV6::V5(v5) => v5.version(),

				            CompatV5ToV6::Compat(compat) => compat.version(),

				        }

				    }

				    pub fn date(&self) -> Option<time::OffsetDateTime> {

				        match self {

				            CompatV5ToV6::V5(v5) => v5.date(),

				            CompatV5ToV6::Compat(compat) => compat.date(),

				        }

				    }

				    pub fn instance_uid(&self) -> Result<Option<uuid::Uuid>> {

				        match self {

				            CompatV5ToV6::V5(v5) => v5.instance_uid(),

				            CompatV5ToV6::Compat(compat) => compat.instance_uid(),

				        }

				    }

				    pub fn indexes(&self) -> Result<Box<dyn Iterator<Item = Result<CompatIndexV5ToV6>> + '_>> {

				        let indexes = match self {

				            CompatV5ToV6::V5(v5) => {

				                Box::new(v5.indexes()?.map(|index| index.map(CompatIndexV5ToV6::from)))

				                    as Box<dyn Iterator<Item = Result<CompatIndexV5ToV6>> + '_>

				            }

				            CompatV5ToV6::Compat(compat) => {

				                Box::new(compat.indexes()?.map(|index| index.map(CompatIndexV5ToV6::from)))

				                    as Box<dyn Iterator<Item = Result<CompatIndexV5ToV6>> + '_>

				            }

				        };

				        Ok(indexes)

				    }

				    pub fn tasks(

				        &mut self,

				    ) -> Result<Box<dyn Iterator<Item = Result<(v6::Task, Option<Box<UpdateFile>>)>> + '_>> {

				        let instance_uid = self.instance_uid().ok().flatten();

				        let keys = self.keys()?.collect::<Result<Vec<_>>>()?;

				        let tasks = match self {

				            CompatV5ToV6::V5(v5) => v5.tasks(),

				            CompatV5ToV6::Compat(compat) => compat.tasks(),

				        };

				        Ok(Box::new(tasks.map(move |task| {

				            task.map(|(task, content_file)| {

				                let mut task_view: v5::tasks::TaskView = task.clone().into();

				                if task_view.status == v5::Status::Processing {

				                    task_view.started_at = None;

				                }

				                let task = v6::Task {

				                    uid: task_view.uid,

				                    index_uid: task_view.index_uid,

				                    status: match task_view.status {

				                        v5::Status::Enqueued => v6::Status::Enqueued,

				                        v5::Status::Processing => v6::Status::Enqueued,

				                        v5::Status::Succeeded => v6::Status::Succeeded,

				                        v5::Status::Failed => v6::Status::Failed,

				                    },

				                    kind: match task.content {

				                        v5::tasks::TaskContent::IndexCreation { primary_key, .. } => {

				                            v6::Kind::IndexCreation { primary_key }

				                        }

				                        v5::tasks::TaskContent::IndexUpdate { primary_key, .. } => {

				                            v6::Kind::IndexUpdate { primary_key }

				                        }

				                        v5::tasks::TaskContent::IndexDeletion { .. } => v6::Kind::IndexDeletion,

				                        v5::tasks::TaskContent::DocumentAddition {

				                            merge_strategy,

				                            allow_index_creation,

				                            primary_key,

				                            documents_count,

				                            ..

				                        } => v6::Kind::DocumentImport {

				                            primary_key,

				                            documents_count: documents_count as u64,

				                            method: match merge_strategy {

				                                v5::tasks::IndexDocumentsMethod::ReplaceDocuments => {

				                                    v6::milli::update::IndexDocumentsMethod::ReplaceDocuments

				                                }

				                                v5::tasks::IndexDocumentsMethod::UpdateDocuments => {

				                                    v6::milli::update::IndexDocumentsMethod::UpdateDocuments

				                                }

				                            },

				                            allow_index_creation,

				                        },

				                        v5::tasks::TaskContent::DocumentDeletion { deletion, .. } => match deletion

				                        {

				                            v5::tasks::DocumentDeletion::Clear => v6::Kind::DocumentClear,

				                            v5::tasks::DocumentDeletion::Ids(documents_ids) => {

				                                v6::Kind::DocumentDeletion { documents_ids }

				                            }

				                        },

				                        v5::tasks::TaskContent::SettingsUpdate {

				                            allow_index_creation,

				                            is_deletion,

				                            settings,

				                            ..

				                        } => v6::Kind::Settings {

				                            is_deletion,

				                            allow_index_creation,

				                            settings: Box::new(settings.into()),

				                        },

				                        v5::tasks::TaskContent::Dump { uid: _ } => {

				                            // in v6 we compute the dump_uid from the started_at processing time

				                            v6::Kind::DumpCreation { keys: keys.clone(), instance_uid }

				                        }

				                    },

				                    canceled_by: None,

				                    details: task_view.details.map(|details| match details {

				                        v5::Details::DocumentAddition { received_documents, indexed_documents } => {

				                            v6::Details::DocumentAdditionOrUpdate {

				                                received_documents: received_documents as u64,

				                                indexed_documents,

				                            }

				                        }

				                        v5::Details::Settings { settings } => {

				                            v6::Details::SettingsUpdate { settings: Box::new(settings.into()) }

				                        }

				                        v5::Details::IndexInfo { primary_key } => {

				                            v6::Details::IndexInfo { primary_key }

				                        }

				                        v5::Details::DocumentDeletion {

				                            received_document_ids,

				                            deleted_documents,

				                        } => v6::Details::DocumentDeletion {

				                            provided_ids: received_document_ids,

				                            deleted_documents,

				                        },

				                        v5::Details::ClearAll { deleted_documents } => {

				                            v6::Details::ClearAll { deleted_documents }

				                        }

				                        v5::Details::Dump { dump_uid } => {

				                            v6::Details::Dump { dump_uid: Some(dump_uid) }

				                        }

				                    }),

				                    error: task_view.error.map(|e| e.into()),

				                    enqueued_at: task_view.enqueued_at,

				                    started_at: task_view.started_at,

				                    finished_at: task_view.finished_at,

				                };

				                (task, content_file)

				            })

				        })))

				    }

				    pub fn keys(&mut self) -> Result<Box<dyn Iterator<Item = Result<v6::Key>> + '_>> {

				        let keys = match self {

				            CompatV5ToV6::V5(v5) => v5.keys()?,

				            CompatV5ToV6::Compat(compat) => compat.keys(),

				        };

				        Ok(Box::new(keys.map(|key| {

				            key.map(|key| v6::Key {

				                description: key.description,

				                name: key.name,

				                uid: key.uid,

				                actions: key.actions.into_iter().map(|action| action.into()).collect(),

				                indexes: key

				                    .indexes

				                    .into_iter()

				                    .map(|index| match index {

				                        v5::StarOr::Star => v6::IndexUidPattern::all(),

				                        v5::StarOr::Other(uid) => v6::IndexUidPattern::new_unchecked(uid.as_str()),

				                    })

				                    .collect(),

				                expires_at: key.expires_at,

				                created_at: key.created_at,

				                updated_at: key.updated_at,

				            })

				        })))

				    }

				    pub fn features(&self) -> Result<Option<v6::RuntimeTogglableFeatures>> {

				        Ok(None)

				    }

				}

				pub enum CompatIndexV5ToV6 {

				    V5(v5::V5IndexReader),

				    Compat(CompatIndexV4ToV5),

				}

				impl From<v5::V5IndexReader> for CompatIndexV5ToV6 {

				    fn from(index_reader: v5::V5IndexReader) -> Self {

				        Self::V5(index_reader)

				    }

				}

				impl From<CompatIndexV4ToV5> for CompatIndexV5ToV6 {

				    fn from(index_reader: CompatIndexV4ToV5) -> Self {

				        Self::Compat(index_reader)

				    }

				}

				impl CompatIndexV5ToV6 {

				    pub fn new_v5(v5: v5::V5IndexReader) -> CompatIndexV5ToV6 {

				        CompatIndexV5ToV6::V5(v5)

				    }

				    pub fn metadata(&self) -> &crate::IndexMetadata {

				        match self {

				            CompatIndexV5ToV6::V5(v5) => v5.metadata(),

				            CompatIndexV5ToV6::Compat(compat) => compat.metadata(),

				        }

				    }

				    pub fn documents(&mut self) -> Result<Box<dyn Iterator<Item = Result<Document>> + '_>> {

				        match self {

				            CompatIndexV5ToV6::V5(v5) => v5

				                .documents()

				                .map(|iter| Box::new(iter) as Box<dyn Iterator<Item = Result<Document>> + '_>),

				            CompatIndexV5ToV6::Compat(compat) => compat

				                .documents()

				                .map(|iter| Box::new(iter) as Box<dyn Iterator<Item = Result<Document>> + '_>),

				        }

				    }

				    pub fn settings(&mut self) -> Result<v6::Settings<v6::Checked>> {

				        match self {

				            CompatIndexV5ToV6::V5(v5) => Ok(v6::Settings::from(v5.settings()?).check()),

				            CompatIndexV5ToV6::Compat(compat) => Ok(v6::Settings::from(compat.settings()?).check()),

				        }

				    }

				}

				impl<T> From<v5::Setting<T>> for v6::Setting<T> {

				    fn from(setting: v5::Setting<T>) -> Self {

				        match setting {

				            v5::Setting::Set(t) => v6::Setting::Set(t),

				            v5::Setting::Reset => v6::Setting::Reset,

				            v5::Setting::NotSet => v6::Setting::NotSet,

				        }

				    }

				}

				impl From<v5::ResponseError> for v6::ResponseError {

				    fn from(error: v5::ResponseError) -> Self {

				        let code = match error.error_code.as_ref() {

				            "index_creation_failed" => v6::Code::IndexCreationFailed,

				            "index_already_exists" => v6::Code::IndexAlreadyExists,

				            "index_not_found" => v6::Code::IndexNotFound,

				            "invalid_index_uid" => v6::Code::InvalidIndexUid,

				            "invalid_min_word_length_for_typo" => v6::Code::InvalidSettingsTypoTolerance,

				            "invalid_state" => v6::Code::InvalidState,

				            "primary_key_inference_failed" => v6::Code::IndexPrimaryKeyNoCandidateFound,

				            "index_primary_key_already_exists" => v6::Code::IndexPrimaryKeyAlreadyExists,

				            "max_fields_limit_exceeded" => v6::Code::MaxFieldsLimitExceeded,

				            "missing_document_id" => v6::Code::MissingDocumentId,

				            "invalid_document_id" => v6::Code::InvalidDocumentId,

				            "invalid_filter" => v6::Code::InvalidSettingsFilterableAttributes,

				            "invalid_sort" => v6::Code::InvalidSettingsSortableAttributes,

				            "bad_parameter" => v6::Code::BadParameter,

				            "bad_request" => v6::Code::BadRequest,

				            "database_size_limit_reached" => v6::Code::DatabaseSizeLimitReached,

				            "document_not_found" => v6::Code::DocumentNotFound,

				            "internal" => v6::Code::Internal,

				            "invalid_geo_field" => v6::Code::InvalidDocumentGeoField,

				            "invalid_ranking_rule" => v6::Code::InvalidSettingsRankingRules,

				            "invalid_store_file" => v6::Code::InvalidStoreFile,

				            "invalid_api_key" => v6::Code::InvalidApiKey,

				            "missing_authorization_header" => v6::Code::MissingAuthorizationHeader,

				            "no_space_left_on_device" => v6::Code::NoSpaceLeftOnDevice,

				            "dump_not_found" => v6::Code::DumpNotFound,

				            "task_not_found" => v6::Code::TaskNotFound,

				            "payload_too_large" => v6::Code::PayloadTooLarge,

				            "unretrievable_document" => v6::Code::UnretrievableDocument,

				            "unsupported_media_type" => v6::Code::UnsupportedMediaType,

				            "dump_already_processing" => v6::Code::DumpAlreadyProcessing,

				            "dump_process_failed" => v6::Code::DumpProcessFailed,

				            "invalid_content_type" => v6::Code::InvalidContentType,

				            "missing_content_type" => v6::Code::MissingContentType,

				            "malformed_payload" => v6::Code::MalformedPayload,

				            "missing_payload" => v6::Code::MissingPayload,

				            "api_key_not_found" => v6::Code::ApiKeyNotFound,

				            "missing_parameter" => v6::Code::BadRequest,

				            "invalid_api_key_actions" => v6::Code::InvalidApiKeyActions,

				            "invalid_api_key_indexes" => v6::Code::InvalidApiKeyIndexes,

				            "invalid_api_key_expires_at" => v6::Code::InvalidApiKeyExpiresAt,

				            "invalid_api_key_description" => v6::Code::InvalidApiKeyDescription,

				            "invalid_api_key_name" => v6::Code::InvalidApiKeyName,

				            "invalid_api_key_uid" => v6::Code::InvalidApiKeyUid,

				            "immutable_field" => v6::Code::BadRequest,

				            "api_key_already_exists" => v6::Code::ApiKeyAlreadyExists,

				            other => {

				                tracing::warn!("Unknown error code {}", other);

				                v6::Code::UnretrievableErrorCode

				            }

				        };

				        v6::ResponseError::from_msg(error.message, code)

				    }

				}

				impl<T> From<v5::Settings<T>> for v6::Settings<v6::Unchecked> {

				    fn from(settings: v5::Settings<T>) -> Self {

				        v6::Settings {

				            displayed_attributes: settings.displayed_attributes.into(),

				            searchable_attributes: settings.searchable_attributes.into(),

				            filterable_attributes: settings.filterable_attributes.into(),

				            sortable_attributes: settings.sortable_attributes.into(),

				            ranking_rules: {

				                match settings.ranking_rules {

				                    v5::settings::Setting::Set(ranking_rules) => {

				                        let mut new_ranking_rules = vec![];

				                        for rule in ranking_rules {

				                            match v6::RankingRuleView::from_str(&rule) {

				                                Ok(new_rule) => {

				                                    new_ranking_rules.push(new_rule);

				                                }

				                                Err(_) => {

				                                    tracing::warn!("Error while importing settings. The ranking rule `{rule}` does not exist anymore.")

				                                }

				                            }

				                        }

				                        v6::Setting::Set(new_ranking_rules)

				                    }

				                    v5::settings::Setting::Reset => v6::Setting::Reset,

				                    v5::settings::Setting::NotSet => v6::Setting::NotSet,

				                }

				            },

				            stop_words: settings.stop_words.into(),

				            non_separator_tokens: v6::Setting::NotSet,

				            separator_tokens: v6::Setting::NotSet,

				            dictionary: v6::Setting::NotSet,

				            synonyms: settings.synonyms.into(),

				            distinct_attribute: settings.distinct_attribute.into(),

				            proximity_precision: v6::Setting::NotSet,

				            typo_tolerance: match settings.typo_tolerance {

				                v5::Setting::Set(typo) => v6::Setting::Set(v6::TypoTolerance {

				                    enabled: typo.enabled.into(),

				                    min_word_size_for_typos: match typo.min_word_size_for_typos {

				                        v5::Setting::Set(t) => v6::Setting::Set(v6::MinWordSizeForTypos {

				                            one_typo: t.one_typo.into(),

				                            two_typos: t.two_typos.into(),

				                        }),

				                        v5::Setting::Reset => v6::Setting::Reset,

				                        v5::Setting::NotSet => v6::Setting::NotSet,

				                    },

				                    disable_on_words: typo.disable_on_words.into(),

				                    disable_on_attributes: typo.disable_on_attributes.into(),

				                }),

				                v5::Setting::Reset => v6::Setting::Reset,

				                v5::Setting::NotSet => v6::Setting::NotSet,

				            },

				            faceting: match settings.faceting {

				                v5::Setting::Set(faceting) => v6::Setting::Set(v6::FacetingSettings {

				                    max_values_per_facet: faceting.max_values_per_facet.into(),

				                    sort_facet_values_by: v6::Setting::NotSet,

				                }),

				                v5::Setting::Reset => v6::Setting::Reset,

				                v5::Setting::NotSet => v6::Setting::NotSet,

				            },

				            pagination: match settings.pagination {

				                v5::Setting::Set(pagination) => v6::Setting::Set(v6::PaginationSettings {

				                    max_total_hits: pagination.max_total_hits.into(),

				                }),

				                v5::Setting::Reset => v6::Setting::Reset,

				                v5::Setting::NotSet => v6::Setting::NotSet,

				            },

				            embedders: v6::Setting::NotSet,

				            _kind: std::marker::PhantomData,

				        }

				    }

				}

				impl From<v5::Action> for v6::Action {

				    fn from(key: v5::Action) -> Self {

				        match key {

				            v5::Action::All => v6::Action::All,

				            v5::Action::Search => v6::Action::Search,

				            v5::Action::DocumentsAll => v6::Action::DocumentsAll,

				            v5::Action::DocumentsAdd => v6::Action::DocumentsAdd,

				            v5::Action::DocumentsGet => v6::Action::DocumentsGet,

				            v5::Action::DocumentsDelete => v6::Action::DocumentsDelete,

				            v5::Action::IndexesAll => v6::Action::IndexesAll,

				            v5::Action::IndexesAdd => v6::Action::IndexesAdd,

				            v5::Action::IndexesGet => v6::Action::IndexesGet,

				            v5::Action::IndexesUpdate => v6::Action::IndexesUpdate,

				            v5::Action::IndexesDelete => v6::Action::IndexesDelete,

				            v5::Action::TasksAll => v6::Action::TasksAll,

				            v5::Action::TasksGet => v6::Action::TasksGet,

				            v5::Action::SettingsAll => v6::Action::SettingsAll,

				            v5::Action::SettingsGet => v6::Action::SettingsGet,

				            v5::Action::SettingsUpdate => v6::Action::SettingsUpdate,

				            v5::Action::StatsAll => v6::Action::StatsAll,

				            v5::Action::StatsGet => v6::Action::StatsGet,

				            v5::Action::MetricsAll => v6::Action::MetricsAll,

				            v5::Action::MetricsGet => v6::Action::MetricsGet,

				            v5::Action::DumpsAll => v6::Action::DumpsAll,

				            v5::Action::DumpsCreate => v6::Action::DumpsCreate,

				            v5::Action::Version => v6::Action::Version,

				            v5::Action::KeysAdd => v6::Action::KeysAdd,

				            v5::Action::KeysGet => v6::Action::KeysGet,

				            v5::Action::KeysUpdate => v6::Action::KeysUpdate,

				            v5::Action::KeysDelete => v6::Action::KeysDelete,

				        }

				    }

				}

				#[cfg(test)]

				pub(crate) mod test {

				    use std::fs::File;

				    use std::io::BufReader;

				    use flate2::bufread::GzDecoder;

				    use meili_snap::insta;

				    use tempfile::TempDir;

				    use super::*;

				    #[test]

				    fn compat_v5_v6() {

				        let dump = File::open("tests/assets/v5.dump").unwrap();

				        let dir = TempDir::new().unwrap();

				        let mut dump = BufReader::new(dump);

				        let gz = GzDecoder::new(&mut dump);

				        let mut archive = tar::Archive::new(gz);

				        archive.unpack(dir.path()).unwrap();

				        let mut dump = v5::V5Reader::open(dir).unwrap().to_v6();

				        // top level infos

				        insta::assert_display_snapshot!(dump.date().unwrap(), @"2022-10-04 15:55:10.344982459 +00:00:00");

				        insta::assert_display_snapshot!(dump.instance_uid().unwrap().unwrap(), @"9e15e977-f2ae-4761-943f-1eaf75fd736d");

				        // tasks

				        let tasks = dump.tasks().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        let (tasks, update_files): (Vec<_>, Vec<_>) = tasks.into_iter().unzip();

				        meili_snap::snapshot_hash!(meili_snap::json_string!(tasks), @"41f91d3a94911b2735ec41b07540df5c");

				        assert_eq!(update_files.len(), 22);

				        assert!(update_files[0].is_none()); // the dump creation

				        assert!(update_files[1].is_some()); // the enqueued document addition

				        assert!(update_files[2..].iter().all(|u| u.is_none())); // everything already processed

				        // keys

				        let keys = dump.keys().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        meili_snap::snapshot_hash!(meili_snap::json_string!(keys), @"c9d2b467fe2fca0b35580d8a999808fb");

				        // indexes

				        let mut indexes = dump.indexes().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        // the index are not ordered in any way by default

				        indexes.sort_by_key(|index| index.metadata().uid.to_string());

				        let mut products = indexes.pop().unwrap();

				        let mut movies = indexes.pop().unwrap();

				        let mut spells = indexes.pop().unwrap();

				        assert!(indexes.is_empty());

				        // products

				        insta::assert_json_snapshot!(products.metadata(), { ".createdAt" => "[now]", ".updatedAt" => "[now]" }, @r###"

				        {

				          "uid": "products",

				          "primaryKey": "sku",

				          "createdAt": "[now]",

				          "updatedAt": "[now]"

				        }

				        "###);

				        insta::assert_json_snapshot!(products.settings().unwrap());

				        let documents = products.documents().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        assert_eq!(documents.len(), 10);

				        meili_snap::snapshot_hash!(format!("{:#?}", documents), @"b01c8371aea4c7171af0d4d846a2bdca");

				        // movies

				        insta::assert_json_snapshot!(movies.metadata(), { ".createdAt" => "[now]", ".updatedAt" => "[now]" }, @r###"

				        {

				          "uid": "movies",

				          "primaryKey": "id",

				          "createdAt": "[now]",

				          "updatedAt": "[now]"

				        }

				        "###);

				        insta::assert_json_snapshot!(movies.settings().unwrap());

				        let documents = movies.documents().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        assert_eq!(documents.len(), 200);

				        meili_snap::snapshot_hash!(format!("{:#?}", documents), @"e962baafd2fbae4cdd14e876053b0c5a");

				        // spells

				        insta::assert_json_snapshot!(spells.metadata(), { ".createdAt" => "[now]", ".updatedAt" => "[now]" }, @r###"

				        {

				          "uid": "dnd_spells",

				          "primaryKey": "index",

				          "createdAt": "[now]",

				          "updatedAt": "[now]"

				        }

				        "###);

				        insta::assert_json_snapshot!(spells.settings().unwrap());

				        let documents = spells.documents().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        assert_eq!(documents.len(), 10);

				        meili_snap::snapshot_hash!(format!("{:#?}", documents), @"235016433dd04262c7f2da01d1e808ce");

				    }

				}

									
										741

dump/src/reader/mod.rs
									
										Normal file
									
												View File
												
				@@ -0,0 +1,741 @@

				use std::fs::File;

				use std::io::{BufReader, Read};

				use flate2::bufread::GzDecoder;

				use serde::Deserialize;

				use tempfile::TempDir;

				use self::compat::v4_to_v5::CompatV4ToV5;

				use self::compat::v5_to_v6::{CompatIndexV5ToV6, CompatV5ToV6};

				use self::v5::V5Reader;

				use self::v6::{V6IndexReader, V6Reader};

				use crate::{Result, Version};

				mod compat;

				mod v1;

				mod v2;

				mod v3;

				mod v4;

				mod v5;

				mod v6;

				pub type Document = serde_json::Map<String, serde_json::Value>;

				pub type UpdateFile = dyn Iterator<Item = Result<Document>>;

				pub enum DumpReader {

				    Current(V6Reader),

				    Compat(CompatV5ToV6),

				}

				impl DumpReader {

				    pub fn open(dump: impl Read) -> Result<DumpReader> {

				        let path = TempDir::new()?;

				        let mut dump = BufReader::new(dump);

				        let gz = GzDecoder::new(&mut dump);

				        let mut archive = tar::Archive::new(gz);

				        archive.unpack(path.path())?;

				        #[derive(Deserialize)]

				        #[serde(rename_all = "camelCase")]

				        struct MetadataVersion {

				            pub dump_version: Version,

				        }

				        let mut meta_file = File::open(path.path().join("metadata.json"))?;

				        let MetadataVersion { dump_version } = serde_json::from_reader(&mut meta_file)?;

				        match dump_version {

				            Version::V1 => {

				                Ok(v1::V1Reader::open(path)?.to_v2().to_v3().to_v4().to_v5().to_v6().into())

				            }

				            Version::V2 => Ok(v2::V2Reader::open(path)?.to_v3().to_v4().to_v5().to_v6().into()),

				            Version::V3 => Ok(v3::V3Reader::open(path)?.to_v4().to_v5().to_v6().into()),

				            Version::V4 => Ok(v4::V4Reader::open(path)?.to_v5().to_v6().into()),

				            Version::V5 => Ok(v5::V5Reader::open(path)?.to_v6().into()),

				            Version::V6 => Ok(v6::V6Reader::open(path)?.into()),

				        }

				    }

				    pub fn version(&self) -> crate::Version {

				        match self {

				            DumpReader::Current(current) => current.version(),

				            DumpReader::Compat(compat) => compat.version(),

				        }

				    }

				    pub fn date(&self) -> Option<time::OffsetDateTime> {

				        match self {

				            DumpReader::Current(current) => current.date(),

				            DumpReader::Compat(compat) => compat.date(),

				        }

				    }

				    pub fn instance_uid(&self) -> Result<Option<uuid::Uuid>> {

				        match self {

				            DumpReader::Current(current) => current.instance_uid(),

				            DumpReader::Compat(compat) => compat.instance_uid(),

				        }

				    }

				    pub fn indexes(&self) -> Result<Box<dyn Iterator<Item = Result<DumpIndexReader>> + '_>> {

				        match self {

				            DumpReader::Current(current) => {

				                let indexes = Box::new(current.indexes()?.map(|res| res.map(DumpIndexReader::from)))

				                    as Box<dyn Iterator<Item = Result<DumpIndexReader>> + '_>;

				                Ok(indexes)

				            }

				            DumpReader::Compat(compat) => {

				                let indexes = Box::new(compat.indexes()?.map(|res| res.map(DumpIndexReader::from)))

				                    as Box<dyn Iterator<Item = Result<DumpIndexReader>> + '_>;

				                Ok(indexes)

				            }

				        }

				    }

				    pub fn tasks(

				        &mut self,

				    ) -> Result<Box<dyn Iterator<Item = Result<(v6::Task, Option<Box<UpdateFile>>)>> + '_>> {

				        match self {

				            DumpReader::Current(current) => Ok(current.tasks()),

				            DumpReader::Compat(compat) => compat.tasks(),

				        }

				    }

				    pub fn keys(&mut self) -> Result<Box<dyn Iterator<Item = Result<v6::Key>> + '_>> {

				        match self {

				            DumpReader::Current(current) => Ok(current.keys()),

				            DumpReader::Compat(compat) => compat.keys(),

				        }

				    }

				    pub fn features(&self) -> Result<Option<v6::RuntimeTogglableFeatures>> {

				        match self {

				            DumpReader::Current(current) => Ok(current.features()),

				            DumpReader::Compat(compat) => compat.features(),

				        }

				    }

				}

				impl From<V6Reader> for DumpReader {

				    fn from(value: V6Reader) -> Self {

				        DumpReader::Current(value)

				    }

				}

				impl From<CompatV5ToV6> for DumpReader {

				    fn from(value: CompatV5ToV6) -> Self {

				        DumpReader::Compat(value)

				    }

				}

				impl From<V5Reader> for DumpReader {

				    fn from(value: V5Reader) -> Self {

				        DumpReader::Compat(value.to_v6())

				    }

				}

				impl From<CompatV4ToV5> for DumpReader {

				    fn from(value: CompatV4ToV5) -> Self {

				        DumpReader::Compat(value.to_v6())

				    }

				}

				pub enum DumpIndexReader {

				    Current(v6::V6IndexReader),

				    Compat(Box<CompatIndexV5ToV6>),

				}

				impl DumpIndexReader {

				    pub fn new_v6(v6: v6::V6IndexReader) -> DumpIndexReader {

				        DumpIndexReader::Current(v6)

				    }

				    pub fn metadata(&self) -> &crate::IndexMetadata {

				        match self {

				            DumpIndexReader::Current(v6) => v6.metadata(),

				            DumpIndexReader::Compat(compat) => compat.metadata(),

				        }

				    }

				    pub fn documents(&mut self) -> Result<Box<dyn Iterator<Item = Result<Document>> + '_>> {

				        match self {

				            DumpIndexReader::Current(v6) => v6

				                .documents()

				                .map(|iter| Box::new(iter) as Box<dyn Iterator<Item = Result<Document>> + '_>),

				            DumpIndexReader::Compat(compat) => compat

				                .documents()

				                .map(|iter| Box::new(iter) as Box<dyn Iterator<Item = Result<Document>> + '_>),

				        }

				    }

				    pub fn settings(&mut self) -> Result<v6::Settings<v6::Checked>> {

				        match self {

				            DumpIndexReader::Current(v6) => v6.settings(),

				            DumpIndexReader::Compat(compat) => compat.settings(),

				        }

				    }

				}

				impl From<V6IndexReader> for DumpIndexReader {

				    fn from(value: V6IndexReader) -> Self {

				        DumpIndexReader::Current(value)

				    }

				}

				impl From<CompatIndexV5ToV6> for DumpIndexReader {

				    fn from(value: CompatIndexV5ToV6) -> Self {

				        DumpIndexReader::Compat(Box::new(value))

				    }

				}

				#[cfg(test)]

				pub(crate) mod test {

				    use std::fs::File;

				    use meili_snap::insta;

				    use super::*;

				    use crate::reader::v6::RuntimeTogglableFeatures;

				    #[test]

				    fn import_dump_v6_experimental() {

				        let dump = File::open("tests/assets/v6-with-experimental.dump").unwrap();

				        let mut dump = DumpReader::open(dump).unwrap();

				        // top level infos

				        insta::assert_display_snapshot!(dump.date().unwrap(), @"2023-07-06 7:10:27.21958 +00:00:00");

				        insta::assert_debug_snapshot!(dump.instance_uid().unwrap(), @"None");

				        // tasks

				        let tasks = dump.tasks().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        let (tasks, update_files): (Vec<_>, Vec<_>) = tasks.into_iter().unzip();

				        meili_snap::snapshot_hash!(meili_snap::json_string!(tasks), @"d45cd8571703e58ae53c7bd7ce3f5c22");

				        assert_eq!(update_files.len(), 2);

				        assert!(update_files[0].is_none()); // the dump creation

				        assert!(update_files[1].is_none()); // the processed document addition

				        // keys

				        let keys = dump.keys().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        meili_snap::snapshot_hash!(meili_snap::json_string!(keys), @"13c2da155e9729c2344688cab29af71d");

				        // indexes

				        let mut indexes = dump.indexes().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        // the index are not ordered in any way by default

				        indexes.sort_by_key(|index| index.metadata().uid.to_string());

				        let mut test = indexes.pop().unwrap();

				        assert!(indexes.is_empty());

				        insta::assert_json_snapshot!(test.metadata(), @r###"

				        {

				          "uid": "test",

				          "primaryKey": "id",

				          "createdAt": "2023-07-06T07:07:41.364694Z",

				          "updatedAt": "2023-07-06T07:07:41.396114Z"

				        }

				        "###);

				        assert_eq!(test.documents().unwrap().count(), 1);

				        assert_eq!(

				            dump.features().unwrap().unwrap(),

				            RuntimeTogglableFeatures { vector_store: true, ..Default::default() }

				        );

				    }

				    #[test]

				    fn import_dump_v5() {

				        let dump = File::open("tests/assets/v5.dump").unwrap();

				        let mut dump = DumpReader::open(dump).unwrap();

				        // top level infos

				        insta::assert_display_snapshot!(dump.date().unwrap(), @"2022-10-04 15:55:10.344982459 +00:00:00");

				        insta::assert_display_snapshot!(dump.instance_uid().unwrap().unwrap(), @"9e15e977-f2ae-4761-943f-1eaf75fd736d");

				        // tasks

				        let tasks = dump.tasks().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        let (tasks, update_files): (Vec<_>, Vec<_>) = tasks.into_iter().unzip();

				        meili_snap::snapshot_hash!(meili_snap::json_string!(tasks), @"41f91d3a94911b2735ec41b07540df5c");

				        assert_eq!(update_files.len(), 22);

				        assert!(update_files[0].is_none()); // the dump creation

				        assert!(update_files[1].is_some()); // the enqueued document addition

				        assert!(update_files[2..].iter().all(|u| u.is_none())); // everything already processed

				        // keys

				        let keys = dump.keys().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        meili_snap::snapshot_hash!(meili_snap::json_string!(keys), @"c9d2b467fe2fca0b35580d8a999808fb");

				        // indexes

				        let mut indexes = dump.indexes().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        // the index are not ordered in any way by default

				        indexes.sort_by_key(|index| index.metadata().uid.to_string());

				        let mut products = indexes.pop().unwrap();

				        let mut movies = indexes.pop().unwrap();

				        let mut spells = indexes.pop().unwrap();

				        assert!(indexes.is_empty());

				        // products

				        insta::assert_json_snapshot!(products.metadata(), @r###"

				        {

				          "uid": "products",

				          "primaryKey": "sku",

				          "createdAt": "2022-10-04T15:51:35.939396731Z",

				          "updatedAt": "2022-10-04T15:55:01.897325373Z"

				        }

				        "###);

				        insta::assert_json_snapshot!(products.settings().unwrap());

				        let documents = products.documents().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        assert_eq!(documents.len(), 10);

				        meili_snap::snapshot_hash!(format!("{:#?}", documents), @"b01c8371aea4c7171af0d4d846a2bdca");

				        // movies

				        insta::assert_json_snapshot!(movies.metadata(), @r###"

				        {

				          "uid": "movies",

				          "primaryKey": "id",

				          "createdAt": "2022-10-04T15:51:35.291992167Z",

				          "updatedAt": "2022-10-04T15:55:10.33561842Z"

				        }

				        "###);

				        insta::assert_json_snapshot!(movies.settings().unwrap());

				        let documents = movies.documents().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        assert_eq!(documents.len(), 200);

				        meili_snap::snapshot_hash!(format!("{:#?}", documents), @"e962baafd2fbae4cdd14e876053b0c5a");

				        // spells

				        insta::assert_json_snapshot!(spells.metadata(), @r###"

				        {

				          "uid": "dnd_spells",

				          "primaryKey": "index",

				          "createdAt": "2022-10-04T15:51:37.381094632Z",

				          "updatedAt": "2022-10-04T15:55:02.394503431Z"

				        }

				        "###);

				        insta::assert_json_snapshot!(spells.settings().unwrap());

				        let documents = spells.documents().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        assert_eq!(documents.len(), 10);

				        meili_snap::snapshot_hash!(format!("{:#?}", documents), @"235016433dd04262c7f2da01d1e808ce");

				        assert_eq!(dump.features().unwrap(), None);

				    }

				    #[test]

				    fn import_dump_v4() {

				        let dump = File::open("tests/assets/v4.dump").unwrap();

				        let mut dump = DumpReader::open(dump).unwrap();

				        // top level infos

				        insta::assert_display_snapshot!(dump.date().unwrap(), @"2022-10-06 12:53:49.131989609 +00:00:00");

				        insta::assert_display_snapshot!(dump.instance_uid().unwrap().unwrap(), @"9e15e977-f2ae-4761-943f-1eaf75fd736d");

				        // tasks

				        let tasks = dump.tasks().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        let (tasks, update_files): (Vec<_>, Vec<_>) = tasks.into_iter().unzip();

				        meili_snap::snapshot_hash!(meili_snap::json_string!(tasks), @"c2445ddd1785528b80f2ba534d3bd00c");

				        assert_eq!(update_files.len(), 10);

				        assert!(update_files[0].is_some()); // the enqueued document addition

				        assert!(update_files[1..].iter().all(|u| u.is_none())); // everything already processed

				        // keys

				        let keys = dump.keys().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        meili_snap::snapshot_hash!(meili_snap::json_string!(keys, { "[].uid" => "[uuid]" }), @"d751713988987e9331980363e24189ce");

				        // indexes

				        let mut indexes = dump.indexes().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        // the index are not ordered in any way by default

				        indexes.sort_by_key(|index| index.metadata().uid.to_string());

				        let mut products = indexes.pop().unwrap();

				        let mut movies = indexes.pop().unwrap();

				        let mut spells = indexes.pop().unwrap();

				        assert!(indexes.is_empty());

				        // products

				        insta::assert_json_snapshot!(products.metadata(), @r###"

				        {

				          "uid": "products",

				          "primaryKey": "sku",

				          "createdAt": "2022-10-06T12:53:39.360187055Z",

				          "updatedAt": "2022-10-06T12:53:40.603035979Z"

				        }

				        "###);

				        insta::assert_json_snapshot!(products.settings().unwrap());

				        let documents = products.documents().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        assert_eq!(documents.len(), 10);

				        meili_snap::snapshot_hash!(format!("{:#?}", documents), @"b01c8371aea4c7171af0d4d846a2bdca");

				        // movies

				        insta::assert_json_snapshot!(movies.metadata(), @r###"

				        {

				          "uid": "movies",

				          "primaryKey": "id",

				          "createdAt": "2022-10-06T12:53:38.710611568Z",

				          "updatedAt": "2022-10-06T12:53:49.785862546Z"

				        }

				        "###);

				        insta::assert_json_snapshot!(movies.settings().unwrap());

				        let documents = movies.documents().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        assert_eq!(documents.len(), 110);

				        meili_snap::snapshot_hash!(format!("{:#?}", documents), @"786022a66ecb992c8a2a60fee070a5ab");

				        // spells

				        insta::assert_json_snapshot!(spells.metadata(), @r###"

				        {

				          "uid": "dnd_spells",

				          "primaryKey": "index",

				          "createdAt": "2022-10-06T12:53:40.831649057Z",

				          "updatedAt": "2022-10-06T12:53:41.116036186Z"

				        }

				        "###);

				        insta::assert_json_snapshot!(spells.settings().unwrap());

				        let documents = spells.documents().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        assert_eq!(documents.len(), 10);

				        meili_snap::snapshot_hash!(format!("{:#?}", documents), @"235016433dd04262c7f2da01d1e808ce");

				    }

				    #[test]

				    fn import_dump_v3() {

				        let dump = File::open("tests/assets/v3.dump").unwrap();

				        let mut dump = DumpReader::open(dump).unwrap();

				        // top level infos

				        insta::assert_display_snapshot!(dump.date().unwrap(), @"2022-10-07 11:39:03.709153554 +00:00:00");

				        assert_eq!(dump.instance_uid().unwrap(), None);

				        // tasks

				        let tasks = dump.tasks().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        let (tasks, update_files): (Vec<_>, Vec<_>) = tasks.into_iter().unzip();

				        meili_snap::snapshot_hash!(meili_snap::json_string!(tasks), @"cd12efd308fe3ed226356a727ab42ed3");

				        assert_eq!(update_files.len(), 10);

				        assert!(update_files[0].is_some()); // the enqueued document addition

				        assert!(update_files[1..].iter().all(|u| u.is_none())); // everything already processed

				        // keys

				        let keys = dump.keys().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        meili_snap::snapshot_hash!(meili_snap::json_string!(keys), @"d751713988987e9331980363e24189ce");

				        // indexes

				        let mut indexes = dump.indexes().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        // the index are not ordered in any way by default

				        indexes.sort_by_key(|index| index.metadata().uid.to_string());

				        let mut products = indexes.pop().unwrap();

				        let mut movies2 = indexes.pop().unwrap();

				        let mut movies = indexes.pop().unwrap();

				        let mut spells = indexes.pop().unwrap();

				        assert!(indexes.is_empty());

				        // products

				        insta::assert_json_snapshot!(products.metadata(), { ".createdAt" => "[now]", ".updatedAt" => "[now]" }, @r###"

				        {

				          "uid": "products",

				          "primaryKey": "sku",

				          "createdAt": "[now]",

				          "updatedAt": "[now]"

				        }

				        "###);

				        insta::assert_json_snapshot!(products.settings().unwrap());

				        let documents = products.documents().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        assert_eq!(documents.len(), 10);

				        meili_snap::snapshot_hash!(format!("{:#?}", documents), @"548284a84de510f71e88e6cdea495cf5");

				        // movies

				        insta::assert_json_snapshot!(movies.metadata(), { ".createdAt" => "[now]", ".updatedAt" => "[now]" }, @r###"

				        {

				          "uid": "movies",

				          "primaryKey": "id",

				          "createdAt": "[now]",

				          "updatedAt": "[now]"

				        }

				        "###);

				        insta::assert_json_snapshot!(movies.settings().unwrap());

				        let documents = movies.documents().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        assert_eq!(documents.len(), 110);

				        meili_snap::snapshot_hash!(format!("{:#?}", documents), @"d153b5a81d8b3cdcbe1dec270b574022");

				        // movies2

				        insta::assert_json_snapshot!(movies2.metadata(), { ".createdAt" => "[now]", ".updatedAt" => "[now]" }, @r###"

				        {

				          "uid": "movies_2",

				          "primaryKey": null,

				          "createdAt": "[now]",

				          "updatedAt": "[now]"

				        }

				        "###);

				        insta::assert_json_snapshot!(movies2.settings().unwrap());

				        let documents = movies2.documents().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        assert_eq!(documents.len(), 0);

				        meili_snap::snapshot_hash!(format!("{:#?}", documents), @"d751713988987e9331980363e24189ce");

				        // spells

				        insta::assert_json_snapshot!(spells.metadata(), { ".createdAt" => "[now]", ".updatedAt" => "[now]" }, @r###"

				        {

				          "uid": "dnd_spells",

				          "primaryKey": "index",

				          "createdAt": "[now]",

				          "updatedAt": "[now]"

				        }

				        "###);

				        insta::assert_json_snapshot!(spells.settings().unwrap());

				        let documents = spells.documents().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        assert_eq!(documents.len(), 10);

				        meili_snap::snapshot_hash!(format!("{:#?}", documents), @"235016433dd04262c7f2da01d1e808ce");

				    }

				    #[test]

				    fn import_dump_v2() {

				        let dump = File::open("tests/assets/v2.dump").unwrap();

				        let mut dump = DumpReader::open(dump).unwrap();

				        // top level infos

				        insta::assert_display_snapshot!(dump.date().unwrap(), @"2022-10-09 20:27:59.904096267 +00:00:00");

				        assert_eq!(dump.instance_uid().unwrap(), None);

				        // tasks

				        let tasks = dump.tasks().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        let (tasks, update_files): (Vec<_>, Vec<_>) = tasks.into_iter().unzip();

				        meili_snap::snapshot_hash!(meili_snap::json_string!(tasks), @"bc616290adfe7d09a624cf6065ca9069");

				        assert_eq!(update_files.len(), 9);

				        assert!(update_files[0].is_some()); // the enqueued document addition

				        assert!(update_files[1..].iter().all(|u| u.is_none())); // everything already processed

				        // keys

				        let keys = dump.keys().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        meili_snap::snapshot_hash!(meili_snap::json_string!(keys), @"d751713988987e9331980363e24189ce");

				        // indexes

				        let mut indexes = dump.indexes().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        // the index are not ordered in any way by default

				        indexes.sort_by_key(|index| index.metadata().uid.to_string());

				        let mut products = indexes.pop().unwrap();

				        let mut movies2 = indexes.pop().unwrap();

				        let mut movies = indexes.pop().unwrap();

				        let mut spells = indexes.pop().unwrap();

				        assert!(indexes.is_empty());

				        // products

				        insta::assert_json_snapshot!(products.metadata(), @r###"

				        {

				          "uid": "products",

				          "primaryKey": "sku",

				          "createdAt": "2022-10-09T20:27:22.688964637Z",

				          "updatedAt": "2022-10-09T20:27:23.951017769Z"

				        }

				        "###);

				        insta::assert_json_snapshot!(products.settings().unwrap());

				        let documents = products.documents().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        assert_eq!(documents.len(), 10);

				        meili_snap::snapshot_hash!(format!("{:#?}", documents), @"548284a84de510f71e88e6cdea495cf5");

				        // movies

				        insta::assert_json_snapshot!(movies.metadata(), @r###"

				        {

				          "uid": "movies",

				          "primaryKey": "id",

				          "createdAt": "2022-10-09T20:27:22.197788495Z",

				          "updatedAt": "2022-10-09T20:28:01.93111053Z"

				        }

				        "###);

				        insta::assert_json_snapshot!(movies.settings().unwrap());

				        let documents = movies.documents().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        assert_eq!(documents.len(), 110);

				        meili_snap::snapshot_hash!(format!("{:#?}", documents), @"d153b5a81d8b3cdcbe1dec270b574022");

				        // movies2

				        insta::assert_json_snapshot!(movies2.metadata(), { ".createdAt" => "[now]", ".updatedAt" => "[now]" }, @r###"

				        {

				          "uid": "movies_2",

				          "primaryKey": null,

				          "createdAt": "[now]",

				          "updatedAt": "[now]"

				        }

				        "###);

				        insta::assert_json_snapshot!(movies2.settings().unwrap());

				        let documents = movies2.documents().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        assert_eq!(documents.len(), 0);

				        meili_snap::snapshot_hash!(format!("{:#?}", documents), @"d751713988987e9331980363e24189ce");

				        // spells

				        insta::assert_json_snapshot!(spells.metadata(), @r###"

				        {

				          "uid": "dnd_spells",

				          "primaryKey": "index",

				          "createdAt": "2022-10-09T20:27:24.242683494Z",

				          "updatedAt": "2022-10-09T20:27:24.312809641Z"

				        }

				        "###);

				        insta::assert_json_snapshot!(spells.settings().unwrap());

				        let documents = spells.documents().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        assert_eq!(documents.len(), 10);

				        meili_snap::snapshot_hash!(format!("{:#?}", documents), @"235016433dd04262c7f2da01d1e808ce");

				    }

				    #[test]

				    fn import_dump_v2_from_meilisearch_v0_22_0_issue_3435() {

				        let dump = File::open("tests/assets/v2-v0.22.0.dump").unwrap();

				        let mut dump = DumpReader::open(dump).unwrap();

				        // top level infos

				        insta::assert_display_snapshot!(dump.date().unwrap(), @"2023-01-30 16:26:09.247261 +00:00:00");

				        assert_eq!(dump.instance_uid().unwrap(), None);

				        // tasks

				        let tasks = dump.tasks().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        let (tasks, update_files): (Vec<_>, Vec<_>) = tasks.into_iter().unzip();

				        meili_snap::snapshot_hash!(meili_snap::json_string!(tasks), @"2db37756d8af1fb7623436b76e8956a6");

				        assert_eq!(update_files.len(), 8);

				        assert!(update_files[0..].iter().all(|u| u.is_none())); // everything already processed

				        // keys

				        let keys = dump.keys().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        meili_snap::snapshot_hash!(meili_snap::json_string!(keys), @"d751713988987e9331980363e24189ce");

				        // indexes

				        let mut indexes = dump.indexes().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        // the index are not ordered in any way by default

				        indexes.sort_by_key(|index| index.metadata().uid.to_string());

				        let mut products = indexes.pop().unwrap();

				        let mut movies = indexes.pop().unwrap();

				        let mut spells = indexes.pop().unwrap();

				        assert!(indexes.is_empty());

				        // products

				        insta::assert_json_snapshot!(products.metadata(), @r###"

				        {

				          "uid": "products",

				          "primaryKey": "sku",

				          "createdAt": "2023-01-30T16:25:56.595257Z",

				          "updatedAt": "2023-01-30T16:25:58.70348Z"

				        }

				        "###);

				        insta::assert_json_snapshot!(products.settings().unwrap());

				        let documents = products.documents().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        assert_eq!(documents.len(), 10);

				        meili_snap::snapshot_hash!(format!("{:#?}", documents), @"548284a84de510f71e88e6cdea495cf5");

				        // movies

				        insta::assert_json_snapshot!(movies.metadata(), @r###"

				        {

				          "uid": "movies",

				          "primaryKey": "id",

				          "createdAt": "2023-01-30T16:25:56.192178Z",

				          "updatedAt": "2023-01-30T16:25:56.455714Z"

				        }

				        "###);

				        insta::assert_json_snapshot!(movies.settings().unwrap());

				        let documents = movies.documents().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        assert_eq!(documents.len(), 10);

				        meili_snap::snapshot_hash!(format!("{:#?}", documents), @"0227598af846e574139ee0b80e03a720");

				        // spells

				        insta::assert_json_snapshot!(spells.metadata(), @r###"

				        {

				          "uid": "dnd_spells",

				          "primaryKey": "index",

				          "createdAt": "2023-01-30T16:25:58.876405Z",

				          "updatedAt": "2023-01-30T16:25:59.079906Z"

				        }

				        "###);

				        insta::assert_json_snapshot!(spells.settings().unwrap());

				        let documents = spells.documents().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        assert_eq!(documents.len(), 10);

				        meili_snap::snapshot_hash!(format!("{:#?}", documents), @"235016433dd04262c7f2da01d1e808ce");

				    }

				    #[test]

				    fn import_dump_v1() {

				        let dump = File::open("tests/assets/v1.dump").unwrap();

				        let mut dump = DumpReader::open(dump).unwrap();

				        // top level infos

				        assert_eq!(dump.date(), None);

				        assert_eq!(dump.instance_uid().unwrap(), None);

				        // tasks

				        let tasks = dump.tasks().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        let (tasks, update_files): (Vec<_>, Vec<_>) = tasks.into_iter().unzip();

				        meili_snap::snapshot_hash!(meili_snap::json_string!(tasks), @"8df6eab075a44b3c1af6b726f9fd9a43");

				        assert_eq!(update_files.len(), 9);

				        assert!(update_files[..].iter().all(|u| u.is_none())); // no update file in dump v1

				        // keys

				        let keys = dump.keys().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        meili_snap::snapshot!(meili_snap::json_string!(keys), @"[]");

				        meili_snap::snapshot_hash!(meili_snap::json_string!(keys), @"d751713988987e9331980363e24189ce");

				        // indexes

				        let mut indexes = dump.indexes().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        // the index are not ordered in any way by default

				        indexes.sort_by_key(|index| index.metadata().uid.to_string());

				        let mut products = indexes.pop().unwrap();

				        let mut movies = indexes.pop().unwrap();

				        let mut spells = indexes.pop().unwrap();

				        assert!(indexes.is_empty());

				        // products

				        insta::assert_json_snapshot!(products.metadata(), @r###"

				        {

				          "uid": "products",

				          "primaryKey": "sku",

				          "createdAt": "2022-10-02T13:23:39.976870431Z",

				          "updatedAt": "2022-10-02T13:27:54.353262482Z"

				        }

				        "###);

				        insta::assert_json_snapshot!(products.settings().unwrap());

				        let documents = products.documents().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        assert_eq!(documents.len(), 10);

				        meili_snap::snapshot_hash!(format!("{:#?}", documents), @"b01c8371aea4c7171af0d4d846a2bdca");

				        // movies

				        insta::assert_json_snapshot!(movies.metadata(), @r###"

				        {

				          "uid": "movies",

				          "primaryKey": "id",

				          "createdAt": "2022-10-02T13:15:29.477512777Z",

				          "updatedAt": "2022-10-02T13:21:12.671204856Z"

				        }

				        "###);

				        insta::assert_json_snapshot!(movies.settings().unwrap());

				        let documents = movies.documents().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        assert_eq!(documents.len(), 10);

				        meili_snap::snapshot_hash!(format!("{:#?}", documents), @"b63dbed5bbc059f3e32bc471ae699bf5");

				        // spells

				        insta::assert_json_snapshot!(spells.metadata(), @r###"

				        {

				          "uid": "dnd_spells",

				          "primaryKey": "index",

				          "createdAt": "2022-10-02T13:38:26.358882984Z",

				          "updatedAt": "2022-10-02T13:38:26.385609433Z"

				        }

				        "###);

				        insta::assert_json_snapshot!(spells.settings().unwrap());

				        let documents = spells.documents().unwrap().collect::<Result<Vec<_>>>().unwrap();

				        assert_eq!(documents.len(), 10);

				        meili_snap::snapshot_hash!(format!("{:#?}", documents), @"aa24c0cfc733d66c396237ad44263bed");

				    }

				}

Compare commits

5592 Commits v0.27.2 ... prototype-

2 .cargo/config.toml Normal file Unescape Escape View File

3 .github/ISSUE_TEMPLATE/bug_report.md vendored Unescape Escape View File

11 .github/ISSUE_TEMPLATE/config.yml vendored Unescape Escape View File

49 .github/ISSUE_TEMPLATE/sprint_issue.md vendored Normal file Unescape Escape View File

12 .github/dependabot.yml vendored Normal file Unescape Escape View File

132 .github/is-latest-release.sh vendored Unescape Escape View File

41 .github/scripts/check-release.sh vendored Normal file Unescape Escape View File

48 .github/scripts/is-latest-release.sh vendored Normal file Unescape Escape View File

20 .github/workflows/README.md vendored Unescape Escape View File

30 .github/workflows/bench-manual.yml vendored Normal file Unescape Escape View File

46 .github/workflows/bench-pr.yml vendored Normal file Unescape Escape View File

25 .github/workflows/bench-push-indexing.yml vendored Normal file Unescape Escape View File

77 .github/workflows/benchmarks-manual.yml vendored Normal file Unescape Escape View File

98 .github/workflows/benchmarks-pr.yml vendored Normal file Unescape Escape View File

79 .github/workflows/benchmarks-push-indexing.yml vendored Normal file Unescape Escape View File

78 .github/workflows/benchmarks-push-search-geo.yml vendored Normal file Unescape Escape View File

78 .github/workflows/benchmarks-push-search-songs.yml vendored Normal file Unescape Escape View File

78 .github/workflows/benchmarks-push-search-wiki.yml vendored Normal file Unescape Escape View File

33 .github/workflows/coverage.yml vendored Unescape Escape View File

24 .github/workflows/dependency-issue.yml vendored Normal file Unescape Escape View File

32 .github/workflows/flaky-tests.yml vendored Normal file Unescape Escape View File

15 .github/workflows/flaky.yml vendored Unescape Escape View File

24 .github/workflows/fuzzer-indexing.yml vendored Normal file Unescape Escape View File

29 .github/workflows/latest-git-tag.yml vendored Normal file Unescape Escape View File

154 .github/workflows/milestone-workflow.yml vendored Normal file Unescape Escape View File

58 .github/workflows/publish-apt-brew-pkg.yml vendored Normal file Unescape Escape View File

173 .github/workflows/publish-binaries.yml vendored Unescape Escape View File

39 .github/workflows/publish-deb-brew-pkg.yml vendored Unescape Escape View File

105 .github/workflows/publish-docker-images.yml vendored Normal file Unescape Escape View File

30 .github/workflows/publish-docker-latest.yml vendored Unescape Escape View File

39 .github/workflows/publish-docker-tag.yml vendored Unescape Escape View File

91 .github/workflows/rust.yml vendored Unescape Escape View File

386 .github/workflows/sdks-tests.yml vendored Normal file Unescape Escape View File

190 .github/workflows/test-suite.yml vendored Normal file Unescape Escape View File

47 .github/workflows/update-cargo-toml-version.yml vendored Normal file Unescape Escape View File

13 .gitignore vendored Unescape Escape View File

5 .rustfmt.toml Normal file Unescape Escape View File

72 CONTRIBUTING.md Unescape Escape View File

4874 Cargo.lock generated View File

62 Cargo.toml Unescape Escape View File

18 Dockerfile Unescape Escape View File

2 LICENSE Unescape Escape View File

19 PROFILING.md Normal file Unescape Escape View File

225 README.md Unescape Escape View File

2 SECURITY.md Unescape Escape View File

BIN assets/demo-dark.gif Normal file View File

BIN assets/demo-light.gif Normal file View File

1393 assets/grafana-dashboard.json Normal file View File

BIN assets/integrations.png Normal file View File

30 assets/meilisearch-logo-dark.svg Normal file Unescape Escape View File

30 assets/meilisearch-logo-light.svg Normal file Unescape Escape View File

6 assets/milli-logo.svg Normal file Unescape Escape View File

BIN assets/profiling-example.png Normal file View File

19 assets/prometheus-basic-scraper.yml Normal file Unescape Escape View File

1 benchmarks/.gitignore vendored Normal file Unescape Escape View File

50 benchmarks/Cargo.toml Normal file Unescape Escape View File

138 benchmarks/README.md Normal file Unescape Escape View File

1347 benchmarks/benches/indexing.rs Normal file View File

122 benchmarks/benches/search_geo.rs Normal file Unescape Escape View File

196 benchmarks/benches/search_songs.rs Normal file Unescape Escape View File

129 benchmarks/benches/search_wiki.rs Normal file Unescape Escape View File

256 benchmarks/benches/utils.rs Normal file Unescape Escape View File

115 benchmarks/build.rs Normal file Unescape Escape View File

38 benchmarks/scripts/compare.sh Executable file Unescape Escape View File

14 benchmarks/scripts/list.sh Executable file Unescape Escape View File

5 benchmarks/src/lib.rs Normal file Unescape Escape View File

6 bors.toml Unescape Escape View File

18 build-info/Cargo.toml Normal file Unescape Escape View File

22 build-info/build.rs Normal file Unescape Escape View File

203 build-info/src/lib.rs Normal file Unescape Escape View File

134 config.toml Normal file Unescape Escape View File

247 download-latest.sh Unescape Escape View File

35 dump/Cargo.toml Normal file Unescape Escape View File

17 dump/README.md Normal file Unescape Escape View File

34 dump/src/error.rs Normal file Unescape Escape View File

493 dump/src/lib.rs Normal file Unescape Escape View File

5 dump/src/reader/compat/mod.rs Normal file Unescape Escape View File

38 dump/src/reader/compat/snapshots/dump__reader__compat__v1_to_v2__test__compat_v1_v2-3.snap Normal file Unescape Escape View File

5592 Commits

v0.27.2 ... prototype-

2

.cargo/config.toml Normal file

View File

3

.github/ISSUE_TEMPLATE/bug_report.md vendored

View File

11

.github/ISSUE_TEMPLATE/config.yml vendored

View File

49

.github/ISSUE_TEMPLATE/sprint_issue.md vendored Normal file

View File

12

.github/dependabot.yml vendored Normal file

View File

132

.github/is-latest-release.sh vendored

View File

41

.github/scripts/check-release.sh vendored Normal file

View File

48

.github/scripts/is-latest-release.sh vendored Normal file

View File

20

.github/workflows/README.md vendored

View File

30

.github/workflows/bench-manual.yml vendored Normal file

View File

46

.github/workflows/bench-pr.yml vendored Normal file

View File

25

.github/workflows/bench-push-indexing.yml vendored Normal file

View File

77

.github/workflows/benchmarks-manual.yml vendored Normal file

View File

98

.github/workflows/benchmarks-pr.yml vendored Normal file

View File

79

.github/workflows/benchmarks-push-indexing.yml vendored Normal file

View File

78

.github/workflows/benchmarks-push-search-geo.yml vendored Normal file

View File

78

.github/workflows/benchmarks-push-search-songs.yml vendored Normal file

View File

78

.github/workflows/benchmarks-push-search-wiki.yml vendored Normal file

View File

33

.github/workflows/coverage.yml vendored

View File

24

.github/workflows/dependency-issue.yml vendored Normal file

View File

32

.github/workflows/flaky-tests.yml vendored Normal file

View File

15

.github/workflows/flaky.yml vendored

View File

24

.github/workflows/fuzzer-indexing.yml vendored Normal file

View File

29

.github/workflows/latest-git-tag.yml vendored Normal file

View File

154

.github/workflows/milestone-workflow.yml vendored Normal file

View File

58

.github/workflows/publish-apt-brew-pkg.yml vendored Normal file

View File

173

.github/workflows/publish-binaries.yml vendored

View File

39

.github/workflows/publish-deb-brew-pkg.yml vendored

View File

105

.github/workflows/publish-docker-images.yml vendored Normal file

View File

30

.github/workflows/publish-docker-latest.yml vendored

View File

39

.github/workflows/publish-docker-tag.yml vendored

View File

91

.github/workflows/rust.yml vendored

View File

386

.github/workflows/sdks-tests.yml vendored Normal file

View File

190

.github/workflows/test-suite.yml vendored Normal file

View File

47

.github/workflows/update-cargo-toml-version.yml vendored Normal file

View File

13

.gitignore vendored

View File

5

.rustfmt.toml Normal file

View File

72

CONTRIBUTING.md

View File

4874

Cargo.lock generated

View File

62

Cargo.toml

View File

18

Dockerfile

View File

2

LICENSE

View File

19

PROFILING.md Normal file

View File

225

README.md

View File

2

SECURITY.md

View File

BIN
assets/demo-dark.gif Normal file

View File

BIN
assets/demo-light.gif Normal file

View File

1393

assets/grafana-dashboard.json Normal file

View File

BIN
assets/integrations.png Normal file

View File

30

assets/meilisearch-logo-dark.svg Normal file

View File

30

assets/meilisearch-logo-light.svg Normal file

View File

6

assets/milli-logo.svg Normal file

View File

BIN
assets/profiling-example.png Normal file

View File

19

assets/prometheus-basic-scraper.yml Normal file

View File

1

benchmarks/.gitignore vendored Normal file

View File

50

benchmarks/Cargo.toml Normal file

View File

138

benchmarks/README.md Normal file

View File

1347

benchmarks/benches/indexing.rs Normal file

View File

122

benchmarks/benches/search_geo.rs Normal file

View File

196

benchmarks/benches/search_songs.rs Normal file

View File

129

benchmarks/benches/search_wiki.rs Normal file

View File

256

benchmarks/benches/utils.rs Normal file

View File

115

benchmarks/build.rs Normal file

View File

38

benchmarks/scripts/compare.sh Executable file

View File

14

benchmarks/scripts/list.sh Executable file

View File

5

benchmarks/src/lib.rs Normal file

View File

6

bors.toml

View File

18

build-info/Cargo.toml Normal file

View File

22

build-info/build.rs Normal file

View File

203

build-info/src/lib.rs Normal file

View File

134

config.toml Normal file

View File

247

download-latest.sh

View File

35

dump/Cargo.toml Normal file

View File

17

dump/README.md Normal file

View File

34

dump/src/error.rs Normal file

View File

493

dump/src/lib.rs Normal file

View File

5

dump/src/reader/compat/mod.rs Normal file

View File

38

dump/src/reader/compat/snapshots/dumpreadercompat__v1_to_v2testcompat_v1_v2-3.snap Normal file

View File

31

dump/src/reader/compat/snapshots/dumpreadercompat__v1_to_v2testcompat_v1_v2-6.snap Normal file

View File