elasticsearch

mirror of https://github.com/elastic/elasticsearch.git synced 2025-04-25 15:47:23 -04:00

Author	SHA1	Message	Date
Nik Everett	0683c90ded	REST tests for normalize agg (#89629 ) This adds a REST test for the normalize pipeline agg so we have backwards compatibility tests for it.	2022-08-26 14:18:46 -04:00
Benjamin Trent	46fc42b817	[ML] Make bucket_count_ks_test aggregation generally available (#88657 ) Initially released in 7.14, bucket_count_ks_test is now generally available.	2022-07-25 13:30:48 -04:00
Benjamin Trent	239d45a019	[ML] make bucket_correlation aggregation generally available (#88655 ) Originally released in 7.14, bucket_correlation is now generally available.	2022-07-21 07:20:09 -04:00
Benjamin Trent	237e345d71	[ML][Docs] fix minimum buckets for change_point agg (#86396 )	2022-05-04 09:37:46 -04:00
James Garside	fca3487395	Updated format parameter description to reference Java decimal format (#86163 )	2022-04-25 20:52:44 +01:00
Benjamin Trent	cf151b53fe	[ML] adds new change_point pipeline aggregation (#83428 ) adds a new `change_point` sibling pipeline aggregation. This aggregation detects a change_point in a multi-bucket aggregation. Example: ``` POST kibana_sample_data_flights/_search { "size": 0, "aggs": { "histo": { "date_histogram": { "field": "timestamp", "fixed_interval": "3h" }, "aggs": { "ticket_price": { "max": { "field": "AvgTicketPrice" } } } }, "changes": { "change_point": { "buckets_path": "histo>ticket_price" } } } } ``` Response ``` { /<snip>/ "aggregations" : { "histo" : { "buckets" : [ /<snip>/ ] }, "changes" : { "bucket" : { "key" : "2022-01-28T23:00:00.000Z", "doc_count" : 48, "ticket_price" : { "value" : 1187.61083984375 } }, "type" : { "distribution_change" : { "p_value" : 0.023753965139433175, "change_point" : 40 } } } } } ```	2022-03-04 07:00:58 -05:00
Lisa Cawley	4fbbcda494	[DOCS] Fix nesting in bucket correlation aggregation (#83816 )	2022-02-11 11:14:11 -08:00
James Rodewig	cf30b54a58	[DOCS] Fix typo in gap_policy's default value for serial differencing aggregation (#80893 ) (#80912 ) Co-authored-by: James Rodewig <40268737+jrodewig@users.noreply.github.com> Co-authored-by: Simon Stücher <stchr@users.noreply.github.com>	2021-11-22 13:43:16 -05:00
James Rodewig	f56a0f4b66	[DOCS] Remove `testenv` annotations from doc snippet tests (#80023 ) Removes `testenv` annotations and related code. These annotations originally let you skip x-pack snippet tests in the docs. However, that's no longer possible. Relates to #79309, #31619	2021-11-05 18:38:50 -04:00
István Zoltán Szabó	1d367abffc	[DOCS] Modifies aggregations title abbreviation to follow convention. (#78252 )	2021-09-23 16:22:27 +02:00
edh-oss	62a471aefe	Update JSON parser and snippets (#77983 ) Related to issue #77823 This does the following: - Updates several asciidoc files that contained code snippets with invalid JSON, most involving unnecessary trailing commas. - Makes the switch from the Groovy JSON parser to the Jackson parser, pursuant to the general goal of eliminating Groovy dependence. - Makes testing of JSON validity at build time more strict. Note that this update still allows backslash escaping for any character. Currently that matters because of the file "docs/reference/ml/anomaly-detection/apis/get-datafeed-stats.asciidoc", specifically this part: "attributes" : { "ml.machine_memory" : "$body.datafeeds.0.node.attributes.ml\.machine_memory", "ml.max_open_jobs" : "512" } It's not clear to me what change, if any, is appropriate there. So, I've left in the escaped period and configured the parser to ignore it for the time being.	2021-09-20 11:08:26 +01:00
Benjamin Trent	30cf4dc8be	[ML] adding new KS test pipeline aggregation (#73334 ) This adds a new pipeline aggregation for calculating Kolmogorov–Smirnov test for a given sample and buckets path. For now, the buckets path resolution needs to be `_count`. But, this may be relaxed in the future. It accepts a parameter `fractions` that indicates the distribution of documents from some other pre-calculated sample. This particular version of the K-S test is Two-sample, meaning, it calculates if the `fractions` and the distribution of `_count` values in the buckets_path are taken from the same distribution. This in combination with the hypothesis alternatives (`less`, `greater`, `two_sided`) and sampling logic (`upper_tail`, `lower_tail`, `uniform`) allow for flexibility and usefulness when comparing two samples and determining the likelihood of them being from the same overall distribution. Usage: ``` POST correlate_latency/_search?size=0&filter_path=aggregations { "aggs": { "buckets": { "terms": { <1> "field": "version", "size": 2 }, "aggs": { "latency_ranges": { "range": { <2> "field": "latency", "ranges": [ { "to": 0.0 }, { "from": 0, "to": 105 }, { "from": 105, "to": 225 }, { "from": 225, "to": 445 }, { "from": 445, "to": 665 }, { "from": 665, "to": 885 }, { "from": 885, "to": 1115 }, { "from": 1115, "to": 1335 }, { "from": 1335, "to": 1555 }, { "from": 1555, "to": 1775 }, { "from": 1775 } ] } }, "ks_test": { <3> "bucket_count_ks_test": { "buckets_path": "latency_ranges>_count", "alternative": ["less", "greater", "two_sided"] } } } } } } ```	2021-06-04 10:04:41 -04:00
Benjamin Trent	8069e9b233	[ML] add new bucket_correlation aggregation with initial count_correlation function (#72133 ) This commit adds a new pipeline aggregation that allows correlation within the aggregation frame work in bucketed values. The initial function is a `count_correlation` function. The purpose of which is to correlate the count in a consistent number of buckets with a pre calculated indicator. The indicator and the aggregated buckets should related to the same metrics with in documents. Example for correlating terms within a `service.version.keyword` with latency percentiles. The percentiles and provided correlation indicator both refer to the same source data where the indicator was previously calculated.: ``` GET apm-7.12.0-transaction-generated/_search { "size": 0, "aggs": { "field_terms": { "terms": { "field": "service.version.keyword", "size": 20 }, "aggs": { "latency_range": { "range": { "field": "transaction.duration.us", "ranges": [<snip>], "keyed": true } }, "correlation": { "bucket_correlation": { "buckets_path": "latency_range>_count", "count_correlation": { "indicator": { "expectations": [<snip>], "doc_count": 20000 } } } } } } } } ```	2021-05-10 12:46:11 -04:00
James Rodewig	693807a6d3	[DOCS] Fix double spaces (#71082 )	2021-03-31 09:57:47 -04:00
István Zoltán Szabó	9a8c6fb66f	[DOCS] Removes beta labels from DFA related docs. (#70808 )	2021-03-26 09:46:41 +01:00
James Rodewig	67288a1e4d	[DOCS] Fix gap policy xref	2021-03-03 09:31:02 -05:00
James Rodewig	e21cab640f	[DOCS] Reformat avg bucket agg reference (#69751 )	2021-03-02 13:44:43 -05:00
RomainGeffraye	fe7afb9d36	[DOCS] Update example for `serial_diff` agg (#69635 )	2021-03-01 08:37:29 -05:00
Lisa Cawley	efa9b095aa	[DOCS] Adds model alias to inference processor and agg (#69576 )	2021-02-24 13:12:39 -08:00
Mike Barretta	12c9ee4d80	Update inference-bucket-aggregation.asciidoc tiny change to properly align the first code example and to add a missing word	2020-12-03 11:48:45 -05:00
James Rodewig	8bc922512c	[DOCS] Redirect moving avg aggregation (#64435 )	2020-10-30 14:12:09 -04:00
James Rodewig	2e9f95aa73	[DOCS] Change agg titles to sentence case (#64425 )	2020-10-30 13:25:21 -04:00
István Zoltán Szabó	6093518f4a	[DOCS] Changes experimental flag to beta in DFA related docs (#63992 )	2020-10-26 17:02:46 +01:00
Benjamin Trent	1084aaf18a	[ML] renames /inference apis to /trained_models (#63097 ) This commit renames all `inference` CRUD APIs to `trained_models`. This aligns with internal terminology, documentation, and use-cases.	2020-10-01 12:13:49 -04:00
Lisa Cawley	ecf9e929ba	[DOCS] Add experimental tag to inference processor and bucket aggregation (#63023 )	2020-09-30 07:20:38 -07:00
István Zoltán Szabó	8da6bba0fc	[DOCS] Adds example to the inference aggregation description (#61290 )	2020-08-19 11:20:42 +02:00
Gilad Gal	8534bd5ce7	Update normalize-aggregation.asciidoc The second method normalizes linearly between 0..100	2020-08-12 22:24:36 +03:00
James Rodewig	74c9e56735	[DOCS] Fix default gap policy for moving fn, moving avg aggs (#60223 ) (#60230 )	2020-07-27 12:32:35 -04:00
James Rodewig	d5b03f668b	[DOCS] Move search sort docs to separate page (#60123 ) Moves the search sort docs from the deprecated 'Request Body Search' page to a new subpage of 'Run a search'. No substantive changes were made to the content.	2020-07-23 12:58:57 -04:00
James Rodewig	2c5d6e9c95	[DOCS] Reformat agg snippets to use two-space indents (#59912 )	2020-07-20 15:08:04 -04:00
István Zoltán Szabó	edccf14478	[DOCS] Adds security privilege info to inference bucket aggregation (#59604 )	2020-07-16 18:02:17 +02:00
David Kyle	7daed3b8af	Pipeline Inference Aggregation (#58193 ) Adds a pipeline aggregation that loads a model and performs inference on the input aggregation results.	2020-07-02 14:33:02 +01:00
andrewjohnson2	a791d6723d	Added standard deviation / variance sampling to extended stats (#49782 ) Per 49554 I added standard deviation sampling and variance sampling to the extended stats interface. Closes #49554 Co-authored-by: Igor Motov <igor@motovs.org>	2020-06-10 15:00:50 -04:00
Tal Levy	79367e43da	Add Normalize Pipeline Aggregation (#56399 ) This aggregation will perform normalizations of metrics for a given series of data in the form of bucket values. The aggregations supports the following normalizations - rescale 0-1 - rescale 0-100 - percentage of sum - mean normalization - z-score normalization - softmax normalization To specify which normalization is to be used, it can be specified in the normalize agg's `normalizer` field. For example: ``` { "normalize": { "buckets_path": <>, "normalizer": "percent" } } ``` Closes #51005.	2020-05-14 13:32:42 -07:00
Ignacio Vera	4e39184c38	Add moving percentiles pipeline aggregation (#55441 ) Similar to what the moving function aggregation does, except merging windows of percentiles sketches together instead of cumulatively merging final metrics	2020-05-12 10:30:52 +02:00
Florian Kelbert	0778c34630	[DOCS] Fix typo in bucket sum aggregation docs (#50431 )	2019-12-20 08:47:24 -05:00
James Rodewig	e43be90e6c	[DOCS] [5 of 5] Change // TESTRESPONSE comments to [source,console-results] (#46449 )	2019-09-06 14:05:36 -04:00
James Rodewig	f5827ba0ae	[DOCS] Replace "// CONSOLE" comments with [source,console] (#46159 )	2019-09-04 12:51:02 -04:00
Zachary Tong	273c35f79c	Add Cumulative Cardinality agg (and Data Science plugin) (#43661 ) This adds a pipeline aggregation that calculates the cumulative cardinality of a field. It does this by iteratively merging in the HLL sketch from consecutive buckets and emitting the cardinality up to that point. This is useful for things like finding the total "new" users that have visited a website (as opposed to "repeat" visitors). This is a Basic+ aggregation and adds a new Data Science plugin to house it and future advanced analytics/data science aggregations.	2019-08-26 10:43:24 -04:00
Nikita Glashenko	ead4eb5209	Add more flexibility to MovingFunction window alignment (#44360 ) Introduce shift field to MovingFunction aggregation. By default, shift = 0. Behavior, in this case, is the same as before. Increasing shift by 1 moves starting window position by 1 to the right. To simply include current bucket to the window, use shift = 1 For center alignment (n/2 values before and after the current bucket), use shift = window / 2 For right alignment (n values after the current bucket), use shift = window.	2019-08-02 15:09:48 -04:00
James Rodewig	ea1adb61c2	[DOCS] Update anchors and links for Elasticsearch API relocation (#44500 )	2019-07-19 09:16:35 -04:00
Zachary Tong	290c8b8256	Force selection of calendar or fixed intervals in date histo agg (#33727 ) The date_histogram accepts an interval which can be either a calendar interval (DST-aware, leap seconds, arbitrary length of months, etc) or fixed interval (strict multiples of SI units). Unfortunately this is inferred by first trying to parse as a calendar interval, then falling back to fixed if that fails. This leads to confusing arrangement where `1d` == calendar, but `2d` == fixed. And if you want a day of fixed time, you have to specify `24h` (e.g. the next smallest unit). This arrangement is very error-prone for users. This PR adds `calendar_interval` and `fixed_interval` parameters to any code that uses intervals (date_histogram, rollup, composite, datafeed, etc). Calendar only accepts calendar intervals, fixed accepts any combination of units (meaning `1d` can be used to specify `24h` in fixed time), and both are mutually exclusive. The old interval behavior is deprecated and will throw a deprecation warning. It is also mutually exclusive with the two new parameters. In the future the old dual-purpose interval will be removed. The change applies to both REST and java clients.	2019-05-06 17:17:11 -04:00
James Rodewig	adf67053f4	[DOCS] Add anchors for Asciidoctor migration (#41648 )	2019-04-30 10:19:09 -04:00
Zachary Tong	6f0f8ab4bc	Remove MovingAverage pipeline aggregation (#39328 ) This was deprecated in 6.4.0 and for the entirety of 7.0. Removed in 8.0	2019-03-19 15:31:05 -04:00
Josh Soref	edb48321ba	[DOCS] Various spelling corrections (#37046 )	2019-01-07 14:44:12 +01:00
João Barbosa	276726aea2	Added keyed response to pipeline percentile aggregations 22302 (#36392 ) Closes #22302	2018-12-14 16:22:54 -05:00
Russ Cam	848847d8c7	[Docs] Section header preceded by blank line (#34340 )	2018-11-08 12:44:13 +01:00
lipsill	b7c0d2830a	[Docs] Remove repeating words (#33087 )	2018-08-28 13:16:43 +02:00
Zachary Tong	df853c49c0	Add a MovingFunction pipeline aggregation, deprecate MovingAvg agg (#29594 ) This pipeline aggregation gives the user the ability to script functions that "move" across a window of data, instead of single data points. It is the scripted version of MovingAvg pipeline agg. Through custom script contexts, we expose a number of convenience methods: - MovingFunctions.max() - MovingFunctions.min() - MovingFunctions.sum() - MovingFunctions.unweightedAvg() - MovingFunctions.linearWeightedAvg() - MovingFunctions.ewma() - MovingFunctions.holt() - MovingFunctions.holtWinters() - MovingFunctions.stdDev() The user can also define any arbitrary logic via their own scripting, or combine with the above methods.	2018-05-16 10:57:00 -04:00
Craig van Tonder	95a13ea01c	Update bucket-sort-aggregation.asciidoc (#28937 ) Added two trailing braces that were missing within the first example.	2018-03-08 15:05:34 +01:00

1 2

95 commits