Skip to content

Added support of timestamp/date/time using curly brackets #297

New issue

Have a question about this project? Sign up for a free GitHub account to open an issue and contact its maintainers and the community.

By clicking “Sign up for GitHub”, you agree to our terms of service and privacy statement. We’ll occasionally send you account related emails.

Already on GitHub? Sign in to your account

Merged
merged 8 commits into from
Jul 21, 2023

Conversation

matthewryanwells
Copy link

@matthewryanwells matthewryanwells commented Jul 13, 2023

Description

Added support for using timestamp/date/time using curly brackets in the following ways to fix failing date filters in PowerBI:

Definition Type
{ts 'YYYY-MM-DD hh:mm:ss[.SSS]'} Timestamp
{timestamp 'YYYY-MM-DD hh:mm:ss[.SSS]'} Timestamp
{d 'YYYY-MM-DD'} Date
{date 'YYYY-MM-DD'} Date
{t 'hh:mm:ss[.SSS]'} Time
{time 'hh:mm:ss[.SSS]'} Time

Issues Resolved

opensearch-project#364

Check List

  • New functionality includes testing.
    • All tests pass, including unit test, integration test and doctest
  • New functionality has been documented.
    • New functionality has javadoc added
    • New functionality has user manual doc added
  • Commits are signed per the DCO using --signoff

By submitting this pull request, I confirm that my contribution is made under the terms of the Apache 2.0 license.
For more information on following Developer Certificate of Origin and signing off your commits, please check here.

@codecov
Copy link

codecov bot commented Jul 13, 2023

Codecov Report

Merging #297 (be1d805) into integ-bracketed-timestamp (4102b58) will not change coverage.
The diff coverage is n/a.

❗ Current head be1d805 differs from pull request most recent head 1657498. Consider uploading reports for the commit 1657498 to get more accurate results

@@                     Coverage Diff                      @@
##             integ-bracketed-timestamp     #297   +/-   ##
============================================================
  Coverage                        97.39%   97.39%           
  Complexity                        4603     4603           
============================================================
  Files                              401      401           
  Lines                            11397    11397           
  Branches                           835      835           
============================================================
  Hits                             11100    11100           
  Misses                             290      290           
  Partials                             7        7           
Flag Coverage Δ
sql-engine 97.39% <ø> (ø)

Flags with carried forward coverage won't be shown. Click here to find out more.

📣 We’re building smart automated test selection to slash your CI/CD build times. Learn more

Copy link

@Yury-Fridlyand Yury-Fridlyand left a comment

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

So far so good. Please add

  1. {dt '...'} and {datetime ...} too
  2. UT for parser where you verify that AST from select date '...' and select {d '...'} are equal (and similar tests).
  3. Doctests

@GumpacG
Copy link

GumpacG commented Jul 14, 2023

Could you add tests for invalid string literals? For example, {ts "hello world"} should throw an error.

@matthewryanwells matthewryanwells force-pushed the dev-bracketed-timestamp branch from 9352312 to cb51853 Compare July 14, 2023 17:50
@matthewryanwells matthewryanwells force-pushed the dev-bracketed-timestamp branch from 60e610b to bbc5aa5 Compare July 19, 2023 16:18
@Yury-Fridlyand
Copy link

Checkstyle fails
I think you need to update integ branch too

@matthewryanwells matthewryanwells force-pushed the dev-bracketed-timestamp branch from 9683a57 to be1d805 Compare July 20, 2023 21:37
@matthewryanwells
Copy link
Author

So far so good. Please add

  1. {dt '...'} and {datetime ...} too
  2. UT for parser where you verify that AST from select date '...' and select {d '...'} are equal (and similar tests).
  3. Doctests
  1. We have decided to not add dt/datetime as PowerBI does not use that datatype
  2. I have added unit tests and also added integration tests comparing the output of non bracketed to bracketed outputs
  3. Add doctests

@matthewryanwells
Copy link
Author

Could you add tests for invalid string literals? For example, {ts "hello world"} should throw an error.

Added tests for invalid inputs

@matthewryanwells
Copy link
Author

Checkstyle fails I think you need to update integ branch too

Fixed checkstyle failures and integ branch has been updated

@matthewryanwells matthewryanwells merged commit a7f0182 into integ-bracketed-timestamp Jul 21, 2023
@matthewryanwells matthewryanwells deleted the dev-bracketed-timestamp branch July 21, 2023 21:23
matthewryanwells added a commit that referenced this pull request Jul 21, 2023
matthewryanwells added a commit that referenced this pull request Jul 26, 2023
* added bracketed time/date/timestamp input, tests, and documentation

Signed-off-by: Matthew Wells <[email protected]>
MitchellGale pushed a commit that referenced this pull request Jul 26, 2023
…-project#1894)

* Added support of timestamp/date/time using curly brackets (#297)

* added bracketed time/date/timestamp input, tests, and documentation

Signed-off-by: Matthew Wells <[email protected]>

* improved failing tests

Signed-off-by: Matthew Wells <[email protected]>

* simplified tests for checking for failure

Signed-off-by: Matthew Wells <[email protected]>

* fixed redundant tests and improved tests that should fail

Signed-off-by: Matthew Wells <[email protected]>

---------

Signed-off-by: Matthew Wells <[email protected]>
matthewryanwells added a commit that referenced this pull request Jul 27, 2023
…-project#1894)

* Added support of timestamp/date/time using curly brackets (#297)

* added bracketed time/date/timestamp input, tests, and documentation

Signed-off-by: Matthew Wells <[email protected]>

* improved failing tests

Signed-off-by: Matthew Wells <[email protected]>

* simplified tests for checking for failure

Signed-off-by: Matthew Wells <[email protected]>

* fixed redundant tests and improved tests that should fail

Signed-off-by: Matthew Wells <[email protected]>

---------

Signed-off-by: Matthew Wells <[email protected]>
(cherry picked from commit 1a7134b)
MitchellGale pushed a commit that referenced this pull request Jul 27, 2023
… (opensearch-project#1897)

Signed-off-by: Yury-Fridlyand <[email protected]>

Added support of timestamp/date/time using curly brackets (opensearch-project#1894)

* Added support of timestamp/date/time using curly brackets (#297)

* added bracketed time/date/timestamp input, tests, and documentation

Signed-off-by: Matthew Wells <[email protected]>

* improved failing tests

Signed-off-by: Matthew Wells <[email protected]>

* simplified tests for checking for failure

Signed-off-by: Matthew Wells <[email protected]>

* fixed redundant tests and improved tests that should fail

Signed-off-by: Matthew Wells <[email protected]>

---------

Signed-off-by: Matthew Wells <[email protected]>

Bump bwc version (opensearch-project#1876)

Signed-off-by: Vamsi Manohar <[email protected]>

Prometheus Query Exemplars (opensearch-project#1782)

Signed-off-by: Vamsi Manohar <[email protected]>

Adding google code changes for core/src/main/java/org/opensearch/sql/analysis  core/src/main/java/org/opensearch/sql/ast  core/src/main/java/org/opensearch/sql/data
 core/src/main/java/org/opensearch/sql/datasource

Signed-off-by: Mitchell Gale <[email protected]>

Disable checkstyle fore core module and enable spotless check for the files being checked.

Signed-off-by: Mitchell Gale <[email protected]>

Added back core checkstyle test in sql-test-workflow.yml

Signed-off-by: Mitchell Gale <[email protected]>

Revert "Prometheus Query Exemplars (opensearch-project#1782)"

This reverts commit 430d7a9.

Revert "Bump bwc version (opensearch-project#1876)"

This reverts commit 501392e.

Revert "Added support of timestamp/date/time using curly brackets (opensearch-project#1894)"

This reverts commit 1a7134b.

Revert "Statically init `typeActionMap` in `OpenSearchExprValueFactory`. (#310) (opensearch-project#1897)"

This reverts commit c8d42a7.

Revert "Fix create_index/create_index_with_IOException issue caused by OpenSearch PR change (opensearch-project#1899)"

This reverts commit 7b932a7.
matthewryanwells added a commit that referenced this pull request Jul 27, 2023
…-project#1894)

* Added support of timestamp/date/time using curly brackets (#297)

* added bracketed time/date/timestamp input, tests, and documentation

Signed-off-by: Matthew Wells <[email protected]>

* improved failing tests

Signed-off-by: Matthew Wells <[email protected]>

* simplified tests for checking for failure

Signed-off-by: Matthew Wells <[email protected]>

* fixed redundant tests and improved tests that should fail

Signed-off-by: Matthew Wells <[email protected]>

---------

Signed-off-by: Matthew Wells <[email protected]>
(cherry picked from commit 1a7134b)
MitchellGale pushed a commit that referenced this pull request Jul 28, 2023
… (opensearch-project#1897)

Signed-off-by: Yury-Fridlyand <[email protected]>

Added support of timestamp/date/time using curly brackets (opensearch-project#1894)

* Added support of timestamp/date/time using curly brackets (#297)

* added bracketed time/date/timestamp input, tests, and documentation

Signed-off-by: Matthew Wells <[email protected]>

* improved failing tests

Signed-off-by: Matthew Wells <[email protected]>

* simplified tests for checking for failure

Signed-off-by: Matthew Wells <[email protected]>

* fixed redundant tests and improved tests that should fail

Signed-off-by: Matthew Wells <[email protected]>

---------

Signed-off-by: Matthew Wells <[email protected]>

Bump bwc version (opensearch-project#1876)

Signed-off-by: Vamsi Manohar <[email protected]>

Prometheus Query Exemplars (opensearch-project#1782)

Signed-off-by: Vamsi Manohar <[email protected]>

Adding google code changes for core/src/main/java/org/opensearch/sql/analysis  core/src/main/java/org/opensearch/sql/ast  core/src/main/java/org/opensearch/sql/data
 core/src/main/java/org/opensearch/sql/datasource

Signed-off-by: Mitchell Gale <[email protected]>

Disable checkstyle fore core module and enable spotless check for the files being checked.

Signed-off-by: Mitchell Gale <[email protected]>

Added back core checkstyle test in sql-test-workflow.yml

Signed-off-by: Mitchell Gale <[email protected]>

Revert "Prometheus Query Exemplars (opensearch-project#1782)"

This reverts commit 430d7a9.

Revert "Bump bwc version (opensearch-project#1876)"

This reverts commit 501392e.

Revert "Added support of timestamp/date/time using curly brackets (opensearch-project#1894)"

This reverts commit 1a7134b.

Revert "Statically init `typeActionMap` in `OpenSearchExprValueFactory`. (#310) (opensearch-project#1897)"

This reverts commit c8d42a7.

Revert "Fix create_index/create_index_with_IOException issue caused by OpenSearch PR change (opensearch-project#1899)"

This reverts commit 7b932a7.

Signed-off-by: Mitchell Gale <[email protected]>
MitchellGale pushed a commit that referenced this pull request Jul 28, 2023
…-project#1894)

* Added support of timestamp/date/time using curly brackets (#297)

* added bracketed time/date/timestamp input, tests, and documentation

Signed-off-by: Matthew Wells <[email protected]>

* improved failing tests

Signed-off-by: Matthew Wells <[email protected]>

* simplified tests for checking for failure

Signed-off-by: Matthew Wells <[email protected]>

* fixed redundant tests and improved tests that should fail

Signed-off-by: Matthew Wells <[email protected]>

---------

Signed-off-by: Matthew Wells <[email protected]>
Signed-off-by: Mitchell Gale <[email protected]>
MitchellGale pushed a commit that referenced this pull request Jul 28, 2023
…-project#1894)

* Added support of timestamp/date/time using curly brackets (#297)

* added bracketed time/date/timestamp input, tests, and documentation

Signed-off-by: Matthew Wells <[email protected]>

* improved failing tests

Signed-off-by: Matthew Wells <[email protected]>

* simplified tests for checking for failure

Signed-off-by: Matthew Wells <[email protected]>

* fixed redundant tests and improved tests that should fail

Signed-off-by: Matthew Wells <[email protected]>

---------

Signed-off-by: Matthew Wells <[email protected]>
Signed-off-by: Mitchell Gale <[email protected]>
MitchellGale pushed a commit that referenced this pull request Jul 31, 2023
… (opensearch-project#1897)

Signed-off-by: Yury-Fridlyand <[email protected]>

Added support of timestamp/date/time using curly brackets (opensearch-project#1894)

* Added support of timestamp/date/time using curly brackets (#297)

* added bracketed time/date/timestamp input, tests, and documentation

Signed-off-by: Matthew Wells <[email protected]>

* improved failing tests

Signed-off-by: Matthew Wells <[email protected]>

* simplified tests for checking for failure

Signed-off-by: Matthew Wells <[email protected]>

* fixed redundant tests and improved tests that should fail

Signed-off-by: Matthew Wells <[email protected]>

---------

Signed-off-by: Matthew Wells <[email protected]>

Bump bwc version (opensearch-project#1876)

Signed-off-by: Vamsi Manohar <[email protected]>

Prometheus Query Exemplars (opensearch-project#1782)

Signed-off-by: Vamsi Manohar <[email protected]>

Adding google code changes for core/src/main/java/org/opensearch/sql/analysis  core/src/main/java/org/opensearch/sql/ast  core/src/main/java/org/opensearch/sql/data
 core/src/main/java/org/opensearch/sql/datasource

Signed-off-by: Mitchell Gale <[email protected]>

Disable checkstyle fore core module and enable spotless check for the files being checked.

Signed-off-by: Mitchell Gale <[email protected]>

Added back core checkstyle test in sql-test-workflow.yml

Signed-off-by: Mitchell Gale <[email protected]>

Revert "Prometheus Query Exemplars (opensearch-project#1782)"

This reverts commit 430d7a9.

Revert "Bump bwc version (opensearch-project#1876)"

This reverts commit 501392e.

Revert "Added support of timestamp/date/time using curly brackets (opensearch-project#1894)"

This reverts commit 1a7134b.

Revert "Statically init `typeActionMap` in `OpenSearchExprValueFactory`. (#310) (opensearch-project#1897)"

This reverts commit c8d42a7.

Revert "Fix create_index/create_index_with_IOException issue caused by OpenSearch PR change (opensearch-project#1899)"

This reverts commit 7b932a7.

Signed-off-by: Mitchell Gale <[email protected]>
Yury-Fridlyand pushed a commit that referenced this pull request Aug 2, 2023
…ets (opensearch-project#1908)

* Added support of timestamp/date/time using curly brackets (opensearch-project#1894)

* Added support of timestamp/date/time using curly brackets (#297)

* added bracketed time/date/timestamp input, tests, and documentation

Signed-off-by: Matthew Wells <[email protected]>

* improved failing tests

Signed-off-by: Matthew Wells <[email protected]>

* simplified tests for checking for failure

Signed-off-by: Matthew Wells <[email protected]>

* fixed redundant tests and improved tests that should fail

Signed-off-by: Matthew Wells <[email protected]>

---------

Signed-off-by: Matthew Wells <[email protected]>
(cherry picked from commit 1a7134b)

* fixed bad cherrypick merge conflict

Signed-off-by: Matthew Wells <[email protected]>

---------

Signed-off-by: Matthew Wells <[email protected]>
andy-k-improving pushed a commit that referenced this pull request Nov 16, 2024
* Implement creation of ip2geo feature (#257)

* Update gradle version to 7.6 (#265)

Signed-off-by: Vijayan Balasubramanian <[email protected]>

* Implement creation of ip2geo feature

* Implementation of ip2geo datasource creation
* Implementation of ip2geo processor creation

Signed-off-by: Heemin Kim <[email protected]>
---------

Signed-off-by: Vijayan Balasubramanian <[email protected]>
Signed-off-by: Heemin Kim <[email protected]>
Co-authored-by: Vijayan Balasubramanian <[email protected]>

* Added unit tests with some refactoring of codes (#271)

* Add Unit tests
* Set cache true for search query
* Remove in memory cache implementation (Two way door decision)
 * Relying on search cache without custom cache
* Renamed datasource state from FAILED to CREATE_FAILED
* Renamed class name from *Helper to *Facade
* Changed updateIntervalInDays to updateInterval
* Changed value type of default update_interval from TimeValue to Long
* Read setting value from cluster settings directly

Signed-off-by: Heemin Kim <[email protected]>

* Sync from main (#280)

* Update gradle version to 7.6 (#265)

Signed-off-by: Vijayan Balasubramanian <[email protected]>

* Exclude lombok generated code from jacoco coverage report (#268)

Signed-off-by: Heemin Kim <[email protected]>

* Make jacoco report to be generated faster in local (#267)

Signed-off-by: Heemin Kim <[email protected]>

* Update dependency org.json:json to v20230227 (#273)

Co-authored-by: mend-for-github-com[bot] <50673670+mend-for-github-com[bot]@users.noreply.github.com>

* Baseline owners and maintainers (#275)

Signed-off-by: Vijayan Balasubramanian <[email protected]>

---------

Signed-off-by: Vijayan Balasubramanian <[email protected]>
Signed-off-by: Heemin Kim <[email protected]>
Co-authored-by: Vijayan Balasubramanian <[email protected]>
Co-authored-by: mend-for-github-com[bot] <50673670+mend-for-github-com[bot]@users.noreply.github.com>

* Add datasource name validation (#281)

Signed-off-by: Heemin Kim <[email protected]>

* Refactoring of code (#282)

1. Change variable name from datasourceName to name
2. Change variable name from id to name
3. Added helper methods in test code

Signed-off-by: Heemin Kim <[email protected]>

* Change field name from md5 to sha256 (#285)

Signed-off-by: Heemin Kim <[email protected]>

* Implement get datasource api (#279)

Signed-off-by: Heemin Kim <[email protected]>

* Update index option (#284)

1. Make geodata index as hidden
2. Make geodata index as read only allow delete after creation is done
3. Refresh datasource index immediately after update

Signed-off-by: Heemin Kim <[email protected]>

* Make some fields in manifest file as mandatory (#289)

Signed-off-by: Heemin Kim <[email protected]>

* Create datasource index explicitly (#283)

Signed-off-by: Heemin Kim <[email protected]>

* Add wrapper class of job scheduler lock service (#290)

Signed-off-by: Heemin Kim <[email protected]>

* Remove all unused client attributes (#293)

Signed-off-by: Heemin Kim <[email protected]>

* Update copyright header (#298)

Signed-off-by: Heemin Kim <[email protected]>

* Run system index handling code with stashed thread context (#297)

Signed-off-by: Heemin Kim <[email protected]>

* Reduce lock duration and renew the lock during update (#299)

Signed-off-by: Heemin Kim <[email protected]>

* Implements delete datasource API (#291)

Signed-off-by: Heemin Kim <[email protected]>

* Set User-Agent in http request (#300)

Signed-off-by: Heemin Kim <[email protected]>

* Implement datasource update API (#292)

Signed-off-by: Heemin Kim <[email protected]>

* Refactoring test code (#302)

Make buildGeoJSONFeatureProcessorConfig method to be more general

Signed-off-by: Heemin Kim <[email protected]>

* Add ip2geo processor integ test for failure case (#303)

Signed-off-by: Heemin Kim <[email protected]>

* Bug fix and refactoring of code (#305)

1. Bugfix: Ingest metadata can be null if there is no processor created
2. Refactoring: Moved private method to another class for better testing support
3. Refactoring: Set some private static final variable as public so that unit test can use it
4. Refactoring: Changed string value to static variable

Signed-off-by: Heemin Kim <[email protected]>

* Add integration test for Ip2GeoProcessor (#306)

Signed-off-by: Heemin Kim <[email protected]>

* Add ConcurrentModificationException (#308)

Signed-off-by: Heemin Kim <[email protected]>

* Add integration test for UpdateDatasource API (#307)

Signed-off-by: Heemin Kim <[email protected]>

* Bug fix on lock management and few performance improvements (#310)

* Release lock before response back to caller for update/delete API
* Release lock in background task for creation API
* Change index settings to improve indexing performance

Signed-off-by: Heemin Kim <[email protected]>

* Change index setting from read_only_allow_delete to write (#311)

read_only_allow_delete does not block write to an index.
The disk-based shard allocator may add and remove this block automatically.
Therefore, use index.blocks.write instead.

Signed-off-by: Heemin Kim <[email protected]>

* Fix bug in get datasource API and improve memory usage (#313)

Signed-off-by: Heemin Kim <[email protected]>

* Change package for Strings.hasText (#314) (#317)

Signed-off-by: Heemin Kim <[email protected]>

* Remove jitter and move index setting from DatasourceFacade to DatasourceExtension (#319)

Signed-off-by: Heemin Kim <[email protected]>

* Do not index blank value and do not enrich null property (#320)

Signed-off-by: Heemin Kim <[email protected]>

* Move index setting keys to constants (#321)

Signed-off-by: Heemin Kim <[email protected]>

* Return null index name for expired data (#322)

Return null index name for expired data so that it can be deleted
by clean up process. Clean up process exclude current index from deleting.
Signed-off-by: Heemin Kim <[email protected]>

* Add new fields in datasource (#325)

Signed-off-by: Heemin Kim <[email protected]>

* Delete index once it is expired (#326)

Signed-off-by: Heemin Kim <[email protected]>

* Add restoring event listener (#328)

In the listener, we trigger a geoip data update

Signed-off-by: Heemin Kim <[email protected]>

* Reverse forcemerge and refresh order (#331)

Otherwise, opensearch does not clear old segment files

Signed-off-by: Heemin Kim <[email protected]>

* Removed parameter and settings (#332)

* Removed first_only parameter
* Removed max_concurrency and batch_size setting

first_only parameter was added as current geoip processor has it.
However, the parameter have no benefit for ip2geo processor as we don't do a sequantial search for array data but use multi search.

max_concurrency and batch_size setting is removed as these are only reveal internal implementation and could be a future blocker to improve performance later.

Signed-off-by: Heemin Kim <[email protected]>

* Add a field in datasource for current index name (#333)

Signed-off-by: Heemin Kim <[email protected]>

* Delete GeoIP data indices after restoring complete (#334)

We don't want to use restored GeoIP data indices. Therefore we
delete the indices once restoring process complete.

When GeoIP metadata index is restored, we create a new GeoIP data index instead.

Signed-off-by: Heemin Kim <[email protected]>

* Use bool query for array form of IPs (#335)

Signed-off-by: Heemin Kim <[email protected]>

* Run update/delete request in a new thread (#337)

This is not to block transport thread

Signed-off-by: Heemin Kim <[email protected]>

* Remove IP2Geo processor validation (#336)

Cannot query index to get data to validate IP2Geo processor.
Will add validation when we decide to store some of data in cluster state metadata.

Signed-off-by: Heemin Kim <[email protected]>

* Acquire lock sychronously (#339)

By acquiring lock asychronously, the remaining part of the code
is being run by transport thread which does not allow blocking code.
We want only single update happen in a node using single thread. However,
it cannot be acheived if I acquire lock asynchronously and pass the listener.

Signed-off-by: Heemin Kim <[email protected]>

* Added a cache to store datasource metadata (#338)

Signed-off-by: Heemin Kim <[email protected]>

* Changed class name and package (#341)

Signed-off-by: Heemin Kim <[email protected]>

* Refactoring of code (#342)

1. Changed class name from Ip2GeoCache to Ip2GeoCachedDao
2. Moved the Ip2GeoCachedDao from cache to dao package

Signed-off-by: Heemin Kim <[email protected]>

* Add geo data cache (#340)

Signed-off-by: Heemin Kim <[email protected]>

* Add cache layer to reduce GeoIp data retrieval latency (#343)

Signed-off-by: Heemin Kim <[email protected]>

* Use _primary in query preference and few changes (#347)

1. Use _primary preference to get datasource metadata so that it can read the latest data. RefreshPolicy.IMMEDIATE won't refresh replica shards immediately according to #346
2. Update datasource metadata index mapping
3. Move batch size from static value to setting

Signed-off-by: Heemin Kim <[email protected]>

* Wait until GeoIP data to be replicated to all data nodes (#348)

Signed-off-by: Heemin Kim <[email protected]>

* Update packages according to a change in OpenSearch core (opensearch-project#354)

* Update packages according to a change in OpenSearch core

Signed-off-by: Heemin Kim <[email protected]>

* Update packages according to a change in OpenSearch core (opensearch-project#353)

Signed-off-by: Heemin Kim <[email protected]>

---------

Signed-off-by: Heemin Kim <[email protected]>

---------

Signed-off-by: Vijayan Balasubramanian <[email protected]>
Signed-off-by: Heemin Kim <[email protected]>
Co-authored-by: Vijayan Balasubramanian <[email protected]>
Co-authored-by: mend-for-github-com[bot] <50673670+mend-for-github-com[bot]@users.noreply.github.com>
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment
Labels
None yet
Projects
None yet
Development

Successfully merging this pull request may close these issues.

6 participants