Performance comparison of multisearch and date range bucket aggregation #63

kaituo · 2020-03-16T05:49:18Z

Performance comparison of multisearch and date range bucket aggregation
Currently, we use multisearch in SearchFeatureDao

kaituo · 2020-03-16T05:49:48Z

running on two benchmarks with multisearch and time range query. date_range query is 2~10 times faster than multisearch. Every time I ran a query, I restarted ec2 nodes to invalidate OS and shard cache.

date_range query:

POST nab_art_daily_jumpsup/_search
 {
   "size": 0,
   "aggs": {
     "range": {
       "date_range": { 
         "field": "timestamp",
         "format": "epoch_millis",
         "ranges": [
           {"from" : 1583391275000, "to": 1583391575000},
           {"from" : 1583392475000, "to": 1583392775000},
           ...
        },
       "aggs": {
         "avg_value": {
           "sum": {
             "field": "value"
           }
         }
       }
     }
   }
 }

multisearch query:

GET /nab_art_daily_jumpsup/_msearch?pretty
{"index" : "nab_art_daily_jumpsup"}
{"query":{"bool":{"must":[{"range":{"timestamp":{"from":1583391275000,"to":1583391575000,"include_lower":true,"include_upper":true,"boost":1}}},{"match_all":{"boost":1}}],"adjust_pure_negative":true,"boost":1}},"aggregations":{"7isajm0Bs-gslJuCOan7":{"sum":{"field":"value"}}}}
{"index" : "nab_art_daily_jumpsup"}
...

Environment:

single node cluster
cluster is not loaded (no other workload besides my queries)

Two benchmarks:

nab_art_daily_jumpsup, sparse data, every 5 minutes one matching document. So each multisearch and date_time query match one doc. I did 3 runs. The following is the recorded time (millisecond)

multisearch	date_time
844	75
973	70
1040	59

server metrics, dense data, every 5 minutes multiple matching documents (e.g., 105). So each multisearch and date_time query match multiple docs. I did 3 runs. The following is the recorded time (millisecond)

multisearch	date_time
663	205
716	181
439	224

Preview does not use all data in the given time range as it is costly. Previously, we sample data by issuing multiple queries on shard 0 data. The purpose of the shard 0 query restriction is to reduce system costs. The nab_art_daily_jumpsup data set has one doc in each interval, and the doc is spread out in 5 shards. Even though we issue 360 queries, we only get 70~80 samples back by querying shard 0. Together with interpolated data points, the preview run misses significant portions of data required to train models (400 is the minimum) and thus returns empty preview results. This PR fixes the issue by removing the shard 0 search restriction. Previously, the preview API issues multiple queries encapsulated in a multisearch request (the request can contain 360 search queries at most). The same result could be obtained via a date range query with multiple range buckets. We show a date range query is 2~10 times faster than a multisearch request (opendistro-for-elasticsearch#63). This PR replaces the multisearch request with a date range query. This PR also removes unused field scriptService in SearchFeatureDao. Testing done: - Previous preview unit tests pass. - Manually verified date range queries results are correctly processed by cross checking intermediate logs. - Manually verified preview results with multisearch and date range implementation are the same. - Manually verified preview don't show empty results with the nab_art_daily_jumpsup data set with the fix

* Fix empty preview result due to insufficient sample Preview does not use all data in the given time range as it is costly. Previously, we sample data by issuing multiple queries on shard 0 data. The purpose of the shard 0 query restriction is to reduce system costs. The nab_art_daily_jumpsup data set has one doc in each interval, and the doc is spread out in 5 shards. Even though we issue 360 queries, we only get 70~80 samples back by querying shard 0. Together with interpolated data points, the preview run misses significant portions of data required to train models (400 is the minimum) and thus returns empty preview results. This PR fixes the issue by removing the shard 0 search restriction. Previously, the preview API issues multiple queries encapsulated in a multisearch request (the request can contain 360 search queries at most). The same result could be obtained via a date range query with multiple range buckets. We show a date range query is 2~10 times faster than a multisearch request (#63). This PR replaces the multisearch request with a date range query. This PR also - removes unused field scriptService in SearchFeatureDao. - fixes rest status during exception - fixes a bug in query generation. We generate aggregation query twice: once with filter query, once separately. Testing done: - Previous preview unit tests pass. - Manually verified date range queries results are correctly processed by cross checking intermediate logs. - Manually verified preview results with multisearch and date range implementation are the same. - Manually verified preview don't show empty results with the nab_art_daily_jumpsup data set with the fix

kaituo closed this as completed Mar 16, 2020

kaituo mentioned this issue Mar 16, 2020

Fix empty preview result due to insufficient sample #65

Merged

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Performance comparison of multisearch and date range bucket aggregation #63

Performance comparison of multisearch and date range bucket aggregation #63

kaituo commented Mar 16, 2020

kaituo commented Mar 16, 2020

Performance comparison of multisearch and date range bucket aggregation #63

Performance comparison of multisearch and date range bucket aggregation #63

Comments

kaituo commented Mar 16, 2020

kaituo commented Mar 16, 2020