elasticsearch

mirror of https://github.com/elastic/elasticsearch.git synced 2025-04-19 04:45:07 -04:00

Author	SHA1	Message	Date
Kathleen DeRusso	e280aa5d50	Revert semantic_text model registry changes (#127075 )	2025-04-18 18:36:33 -04:00
James Baiera	7b89f4d4a6	Add ability to redirect ingestion failures on data streams to a failure store (#126973 ) Removes the feature flags and guards that prevent the new failure store functionality from operating in production runtimes.	2025-04-18 16:33:03 -04:00
Dianna Hohensee	72b4ed255b	Add to allocation architecture guide (#125328 ) How master and data nodes communicate about shard allocation	2025-04-18 14:56:27 -04:00
Joe Gallo	b46bee4e47	Correctly handle non-integers in nested paths in the remove processor (#127006 )	2025-04-18 11:46:54 -04:00
Lorenzo Dematté	69f6520b0c	[Entitlements] Validation checks on paths (#126852 ) With this PR we restrict the paths we allow access to, forbidding plugins to specify/request entitlements for reading or writing to specific protected directories. I added this validation to EntitlementInitialization, as I wanted to fail fast and this is the earliest occurrence where we have all we need: PathLookup to resolve relative paths, policies (for plugins, server, agents) and the Paths for the specific directories we want to protect. Relates to ES-10918	2025-04-18 15:36:07 +02:00
Lorenzo Dematté	b6c9584c28	[Entitlements] Add missing outbound_network entitlement to x-pack-core (#126992 ) Add missing outbound_network entitlement to x-pack-core Closes #127003	2025-04-18 10:19:51 +02:00
elasticsearchmachine	36af046441	Merge patch/serverless-fix into main	2025-04-18 04:30:44 +00:00
Brian Seeders	af6dac5c05	Revert "Forward port release notes for v8.17.5 (#127024 )" This reverts commit `66b504a881`.	2025-04-17 16:16:21 -04:00
elasticsearchmachine	66b504a881	Forward port release notes for v8.17.5 (#127024 )	2025-04-17 16:15:42 -04:00
Brian Seeders	2a243d8492	Revert #126441 Add flow-control and remove auto-read in netty4 HTTP pipeline (#127030 ) * Revert "Release buffers in netty test (#126744)" This reverts commit `f9f3defe92`. * Revert "Add flow-control and remove auto-read in netty4 HTTP pipeline (#126441)" This reverts commit `c8805b85d2`.	2025-04-17 12:37:26 -07:00
David Turner	7e62862eab	Clarify queues in thread pool settings (#127027 ) The docs about the queue in a `fixed` pool are a little awkwardly worded, and there is no mention of the queue in a `scaling` pool at all. This commit cleans this area up.	2025-04-17 19:58:02 +01:00
Liam Thompson	b6c9b9b54d	[DOCS] Update URLs for ESQL Kibana generated docs (#127011 )	2025-04-17 18:25:24 +02:00
Samiul Monir	afb83b7551	Updating text_similarity_reranker documentation (#127004 ) * updating documentation to remove duplicate and redundant wording from 9.x * Update links to rerank model landing page --------- Co-authored-by: Liam Thompson <32779855+leemthompo@users.noreply.github.com>	2025-04-17 11:54:19 -04:00
Luca Cavanna	f274ab7402	Remove empty results before merging (#126770 ) We addressed the empty top docs issue with #126385 specifically for scenarios where empty top docs don't go through the wire. Yet they may be serialized from data node back to the coord node, in which case they will no longer be equal to Lucene#EMPTY_TOP_DOCS. This commit expands the existing filtering of empty top docs to include also those that did go through serialization. Closes #126742	2025-04-17 17:36:20 +02:00
Kathleen DeRusso	a72883e8e3	Default new semantic_text fields to use BBQ when models are compatible (#126629 ) * Default new semantic_text fields to use BBQ when models are compatible * Update docs/changelog/126629.yaml * Gate default BBQ by IndexVersion * Cleanup from PR feedback * PR feedback * Fix test * Fix test * PR feedback * Update test to test correct options * Hack alert: Fix issue where mapper service was always being created with current index version	2025-04-17 08:25:10 -04:00
Nick Tindall	270ca0a80a	Add thread pool utilisation metric (#120363 ) There are existing metrics for the active number of threads, but it seems tricky to go from those to a "utilisation" number because all the pools have different sizes. This commit adds `es.thread_pool.{name}.threads.utilization.current` which will be published by all `TaskExecutionTimeTrackingEsThreadPoolExecutor` thread pools (where `EsExecutors.TaskTrackingConfig#trackExecutionTime` is true). The metric is a double gauge indicating what fraction (in [0.0, 1.0]) of the maximum possible execution time was utilised over the polling interval. It's calculated as actualTaskExecutionTime / maximumTaskExecutionTime, so effectively a "mean" value. The metric interval is 60s so brief spikes won't be apparent in the measure, but the initial goal is to use it to detect hot-spotting so the 60s average will probably suffice. Relates ES-10530	2025-04-17 11:49:30 +10:00
Tim Vernum	e53d3ff64b	Update docs to reflect removal of TLSv1.1 (#126892 ) In ES9 and later, we do not enable TLSv1.1 by default, even if the JDK supports it. This updates the docs accordingly. Relates: #121731	2025-04-17 10:15:29 +10:00
Julio	d19b525eb1	Temporarily bypass competitive iteration for filters aggregation (#12… (#126962 ) * Temporarily bypass competitive iteration for filters aggregation (#126956) * Bump versions after 9.0.0 release * fix merge conflict * Remove 8.16 from branches.json * Bring version-bump related changes from main * [bwc] Add bugfix3 project (#126880) * Sync version bump changes from main again --------- Co-authored-by: Benjamin Trent <ben.w.trent@gmail.com> Co-authored-by: elasticsearchmachine <infra-root+elasticsearchmachine@elastic.co> Co-authored-by: elasticsearchmachine <58790826+elasticsearchmachine@users.noreply.github.com> Co-authored-by: Brian Seeders <brian.seeders@elastic.co>	2025-04-16 18:10:01 -06:00
Benjamin Trent	b1f766258b	Temporarily bypass competitive iteration for filters aggregation (#126956 )	2025-04-16 23:08:17 +02:00
Samiul Monir	2e1101cf5e	Updating text_similarity_reranker documentation (#126175 ) * Updating text_similarity_reranker documentation * Updating docs to include urls * remove extra THE from the text --------- Co-authored-by: Elastic Machine <elasticmachine@users.noreply.github.com>	2025-04-16 17:05:30 -04:00
Ryan Ernst	a813949c34	Fix uniquify to handle multiple successive duplicates (#126889 ) CollectionUtils.uniquify is based on C++ std::unique. However, C++ iterators are not quite the same as Java iterators. In particular, advancing them only allows grabbing the value once. This commit reworks uniquify to be based on list indices instead of iterators. closes #126883	2025-04-16 21:00:27 +02:00
Jonathan Buttner	7a0f63c1a0	[ML] Refactor inference request executor to leverage scheduled execution (#126858 ) * Using threadpool schedule and fixing tests * Update docs/changelog/126858.yaml * Clean up * change log	2025-04-16 14:14:02 -04:00
Jonathan Buttner	e42c118ec6	[ML] Adding missing onFailure call for Inference API start model request (#126930 ) * Adding missing onFailure call * Update docs/changelog/126930.yaml	2025-04-16 14:07:13 -04:00
Nik Everett	128144dd6d	ESQL: Add `documents_found` and `values_loaded` (#125631 ) This adds `documents_found` and `values_loaded` to the to the ESQL response: ```json { "took" : 194, "is_partial" : false, "documents_found" : 100000, "values_loaded" : 200000, "columns" : [ { "name" : "a", "type" : "long" }, { "name" : "b", "type" : "long" } ], "values" : [[10, 1]] } ``` These are cheap enough to collect that we can do it for every query and return it with every response. It's small, but it still gives you a reasonable sense of how much work Elasticsearch had to go through to perform the query. I've also added these two fields to the driver profile and task status: ```json "drivers" : [ { "description" : "data", "cluster_name" : "runTask", "node_name" : "runTask-0", "start_millis" : 1742923173077, "stop_millis" : 1742923173087, "took_nanos" : 9557014, "cpu_nanos" : 9091340, "documents_found" : 5, <---- THESE "values_loaded" : 15, <---- THESE "iterations" : 6, ... ``` These are at a high level and should be easy to reason about. We'd like to extract this into a "show me how difficult this running query is" API one day. But today, just plumbing it into the debugging output is good. Any `Operator` can claim to "find documents" or "load values" by overriding a method on its `Operator.Status` implementation: ```java /** * The number of documents found by this operator. Most operators * don't find documents and will return {@code 0} here. / default long documentsFound() { return 0; } /* * The number of values loaded by this operator. Most operators * don't load values and will return {@code 0} here. / default long valuesLoaded() { return 0; } ``` In this PR all of the `LuceneOperator`s declare that each `position` they emit is a "document found" and the `ValuesSourceValuesSourceReaderOperator` says each value it makes is a "value loaded". That's pretty pretty much true. The `LuceneCountOperator` and `LuceneMinMaxOperator` sort of pretend that the count/min/max that they emit is a "document" - but that's good enough to give you a sense of what's going on. It's like* document.	2025-04-16 17:15:25 +02:00
Lorenzo Dematté	115062c643	Fix vec_caps to test for OS support too (on x64) (#126911 ) On x64, we are testing if we support vector capabilities (1 = "basic" = AVX2, 2 = "advanced" = AVX-512) in order to enable and choose a native implementation for some vector functions, using CPUID. However, under some circumstances, this is not sufficient: the OS on which we are running also needs to support AVX/AVX2 etc; basically, it needs to acknowledge it knows about the additional register and that it is able to handle them e.g. in context switches. To do that we need to a) test if the CPU has xsave feature and b) use the xgetbv to test if the OS set it (declaring it supports AVX/AVX2/etc). In most cases this is not needed, as all modern OSes do that, but for some virtualized situations (hypervisors, emulators, etc.) all the component along the chain must support it, and in some cases this is not a given. This PR introduces a change to the x64 version of vec_caps to check for OS support too, and a warning on the Java side in case the CPU supports vector capabilities but those are not enabled at OS level. Tested by passing noxsave to my linux box kernel boot options, and ensuring that the avx flags "disappear" from /proc/cpuinfo, and we fall back to the "no native vector" case. Fixes #126809	2025-04-16 16:06:46 +02:00
Luca Cavanna	df83e881f9	Cancel expired async search task when a remote returns its results (#126583 ) A while ago we enabled using ccs_minimize_roundtrips in async search. This makes it possible for users of async search to send a single search request per remote cluster, and minimize the impact of network latency. With non minimized roundtrips, we have pretty recurring cancellation checks: as part of the execution, we detect that a task expired whenever each shard comes back with its results. In a scenario where the coord node does not hold data, or only remote data is targeted by an async search, we have much less chance of detecting cancellation if roundtrips are minimized. The local coordinator would do nothing other than waiting for the minimized results from each remote cluster. One scenario where we can check for cancellation is when each cluster comes back with its full set of results. This commit adds such check, plus some testing for async search cancellation with minimized roundtrips.	2025-04-16 14:21:59 +02:00
Niels Bauman	5383f0fcdf	Fix `PolicyStepsRegistry` cache concurrency issue (#126840 ) The following order of events was possible: - An ILM policy update cleared `cachedSteps` - ILM retrieves the step definition for an index, this populates `cachedSteps` with the outdated policy - The updated policy is put in `lifecyclePolicyMap` Any subsequent cache retrievals will see the old step definition. By clearing `cachedSteps` _after_ we update `lifecyclePolicyMap`, we ensure eventual consistency between the policy and the cache. Fixes #118406	2025-04-16 13:58:12 +02:00
Liam Thompson	92148cfde3	[DOCS] Update esql-lookup-join.md to mention index mode requirement (#126901 ) * Update esql-lookup-join.md to mention index mode requirement * fix 8.x page mapping metadata	2025-04-16 12:15:45 +02:00
Carson Ip	5860ccb113	[otel-data] Bump plugin version to release _metric_names_hash changes (#126850 ) Bump otel-data plugin version as #120952 missed the bump.	2025-04-16 10:27:19 +01:00
Paul Tavares	ad0c215369	[Security Solution] Add `read` index privileges to `kibana_system` role for Microsoft Defender integration indexes (#126803 ) adds read privilege to the kibana_system role for indexes associated with the Microsoft Defender Integrations. Changes are necessary in order to support Security Solution bi-directional response actions	2025-04-15 15:42:16 -04:00
Ryan Ernst	6174acdc39	Workaround max name limit imposed by Jackson 2.17 (#126806 ) In Jackson 2.15 a maximum string length of 50k characters was introduced. We worked around that by override the length to max int on all parsers created by xcontent. Jackson 2.17 introduced a similar limit on field names. This commit mimics the workaround for string length by overriding the max name length to be unlimited. relates #58952	2025-04-15 11:40:27 -07:00
Liam Thompson	9ca38f93e8	Re-fix elasticsearch highlights for 9.0 (#126859 )	2025-04-15 18:38:05 +02:00
Luigi Dell'Aquila	de42ba37e0	ES\|QL: fix join masking eval (#126614 )	2025-04-15 18:07:05 +02:00
elasticsearchmachine	e030e44881	Prune changelogs after 8.17.5 release	2025-04-15 14:49:20 +00:00
Svilen Mihaylov	02f9af732e	Add multi_match function #121525 (#125062 ) Implement multi_match function for ESQL. Its currently available on snapshot builds pending refinement of the syntax.	2025-04-15 09:38:08 -04:00
Francisco Fernández Castaño	39670d9477	Add IndexingPressureMonitor to monitor large indexing operations (#126372 ) Relates ES-11063	2025-04-15 13:23:18 +02:00
Richard Dennehy	9e3476ef99	permit at+jwt typ header value in jwt access tokens (#126687 ) * permit at+jwt typ header value in jwt access tokens * Update docs/changelog/126687.yaml * address review comments * [CI] Auto commit changes from spotless * update Type Validator tests for parser ignoring case --------- Co-authored-by: elasticsearchmachine <infra-root+elasticsearchmachine@elastic.co>	2025-04-15 11:08:30 +01:00
Liam Thompson	7de46e9897	[DOCS] Update es-connectors-salesforce.md (#126828 ) * [DOCS] Update es-connectors-salesforce.md 9.x equivalent of https://github.com/elastic/elasticsearch/pull/126791 * Reformat known issues section	2025-04-15 11:47:36 +02:00
Liam Thompson	b326ebb1dd	Fix 8.x page mapping for ES release notes (#126820 )	2025-04-15 10:26:29 +02:00
Ignacio Vera	bcee0af23c	Return float[] instead of List<Double> in valueFetcher (#126702 ) We are currently having to hold in heap big list of Double objects which can take big amounts of heap. With this change we can reduce the heap usage by 7x.	2025-04-15 07:03:55 +02:00
Nick Tindall	dfaf3de96e	Allow float settings to be configured with other settings as default (#126751 ) Relates ES-11367	2025-04-15 13:41:01 +10:00
Dan Rubinstein	b917d9a1e0	Revert endpoint creation validation for ELSER and E5 (#126792 ) * Revert endpoint creation validation for ELSER and E5 * Update docs/changelog/126792.yaml * Revert start model deployment being in TransportPutInferenceModelAction --------- Co-authored-by: Elastic Machine <elasticmachine@users.noreply.github.com>	2025-04-14 17:00:33 -04:00
Ryan Ernst	b47bd3adc7	Use terminal reader in keystore add command (#126729 ) When reading a string value from stdin the keystore add command currently looks directly at stdin. However, stdin may also be consumed while reading the keystore password. This commit changes the add command to use the reader from the termainl instead of looking at stdin directly. closes #98115	2025-04-14 12:55:56 -07:00
Mike Pellegrini	85713f78e0	Semantic Text Chunking Indexing Pressure (#125517 ) We have observed many OOMs due to the memory required to inject chunked inference results for semantic_text fields. This PR uses coordinating indexing pressure to account for this memory usage. When indexing pressure memory usage exceeds the threshold set by indexing_pressure.memory.limit, chunked inference result injection will be suspended to prevent OOMs.	2025-04-14 15:55:37 -04:00
Brian Seeders	01a8a9b63b	[docs] Fix 9.0.0 release notes issues	2025-04-14 14:34:31 -04:00
Benjamin Trent	d7a547597e	Fix bbq quantization algorithm but for differently distributed components (#126778 ) We had a silly bug in quantizing vectors in bbq where we were scaling the initial quantile optimization parameters incorrectly given the vector component distribution. In distributions where this has a major impact, the recall results were abysmal and rendered the quantization technique useless. In modern, well distributed components, this change is almost a no-op.	2025-04-15 04:21:51 +10:00
Charlotte Hoblik	8cb449386b	remove redirects.yml (#126774 )	2025-04-14 14:10:49 +02:00
Ignacio Vera	ffdfcec334	Upgrade to Lucene 10.2.0 (#126594 ) This commit upgrade Elasticsearch to lucene 10.2.0	2025-04-14 13:50:52 +02:00
Kofi B	08beb534ef	[DOCS] Added sort order explanation (#125182 ) * Added explanation of sort order and default behavior * Update docs/reference/elasticsearch/rest-apis/sort-search-results.md Co-authored-by: Liam Thompson <32779855+leemthompo@users.noreply.github.com> --------- Co-authored-by: George Wallace <georgewallace@users.noreply.github.com> Co-authored-by: Liam Thompson <32779855+leemthompo@users.noreply.github.com>	2025-04-14 10:28:03 +02:00
Craig Taverner	ec495e9f0b	Make LOOKUP JOIN docs examples fully tested (#126622 ) The current LOOKUP JOIN docs include examples that are not tested by the ES\|QL tests, unlike most other examples in the documentation. This PR fixes that, changing two examples to use existing tests, and adding a new csv-spec file for the remaining four examples. These four are not required to show results, so the tests have empty data and do not require any results. This means we are testing only the syntax (parsing and semantic analysis), which is sufficient for the docs.	2025-04-14 09:57:58 +02:00

1 2 3 4 5 ...

18045 commits