Primary caching 14: don't bake `LatestAt(T-1)` results into low-level range queries #4793

teh-cmc · 2024-01-12T09:36:20Z

Our low-level range APIs used to bake the latest-at results at range.min - 1 into the range results, which is a big problem in a multi tenant setting because range(1, 10) vs. latestat(1) + range(2, 10) are two completely different things.

Side-effect: a plot with a window of len 1 now behaves as expected:

24-01-12_10.12.19.patched.mp4

Part of the primary caching series of PR (index search, joins, deserialization):

Checklist

I have read and agree to Contributor Guide and the Code of Conduct
I've included a screenshot or gif (if applicable)
I have tested the web demo (if applicable):
- Using newly built examples: app.rerun.io
- Using examples from latest main build: app.rerun.io
- Using full set of examples from nightly build: app.rerun.io
The PR title and labels are set such as to maximize their usefulness for the next release's CHANGELOG

teh-cmc · 2024-01-12T09:36:58Z

crates/re_viewer/src/ui/visible_history.rs

) _99% grunt work, the only somewhat interesting thing happens in `query_archetype`_ Our query model always operates with two distinct timestamps: the timestamp you're querying for (`query_time`) vs. the timestamp of the data you get back (`data_time`). This is the result of our latest-at semantics: a query for a point at time `10` can return a point at time `2`. This is important to know when caching the data: a query at time `4` and a query at time `8` that both return the data at time `2` must share the same single entry or the memory budget would explode. This PR just updates all existing latest-at APIs so they return the data time in their response. This was already the case for range APIs. Note that in the case of `query_archetype`, which is a compound API that emits multiple queries, the data time of the final result is the most recent data time among all of its components. A follow-up PR will use the data time to deduplicate entries in the latest-at cache. --- Part of the primary caching series of PR (index search, joins, deserialization): - #4592 - #4593 - #4659 - #4680 - #4681 - #4698 - #4711 - #4712 - #4721 - #4726 - #4773 - #4784 - #4785 - #4793 - #4800

Wumpf

makes a lot more sense!

crates/re_query/tests/archetype_range_tests.rs

crates/re_data_store/tests/data_store.rs

…ation (#4712) Introduces the notion of cache deduplication: given a query at time `4` and a query at time `8` that both returns data at time `2`, they must share a single cache entry. I.e. starting with this PR, scrubbing through the OPF example will not result if more cache memory being used. --- Part of the primary caching series of PR (index search, joins, deserialization): - #4592 - #4593 - #4659 - #4680 - #4681 - #4698 - #4711 - #4712 - #4721 - #4726 - #4773 - #4784 - #4785 - #4793 - #4800

Introduces a dedicated cache bucket for timeless data and properly forwards the information through all APIs downstream. --- Part of the primary caching series of PR (index search, joins, deserialization): - #4592 - #4593 - #4659 - #4680 - #4681 - #4698 - #4711 - #4712 - #4721 - #4726 - #4773 - #4784 - #4785 - #4793 - #4800

This implements cache invalidation via a `StoreSubscriber`. We keep track of the timestamps to invalidate in the `StoreSubscriber`, but we only do the actual removal of components at query time. This is similar to how we handle bucket sorting in the main store: doing it at query time has the benefit that the frame time effectively behaves as natural micro-batching mechanism that vastly improves performance. --- Part of the primary caching series of PR (index search, joins, deserialization): - #4592 - #4593 - #4659 - #4680 - #4681 - #4698 - #4711 - #4712 - #4721 - #4726 - #4773 - #4784 - #4785 - #4793 - #4800

) The primary cache now tracks memory statistics and display them in the memory panel. This immediately highlights a very stupid thing that the cache does: missing optional components that have been turned into streams of default values by the `ArchetypeView` are materialized as such :man_facepalming: - #4779 https://github.com/rerun-io/rerun/assets/2910679/876b264a-3f77-4d91-934e-aa8897bb32fe - Fixes #4730 --- Part of the primary caching series of PR (index search, joins, deserialization): - #4592 - #4593 - #4659 - #4680 - #4681 - #4698 - #4711 - #4712 - #4721 - #4726 - #4773 - #4784 - #4785 - #4793 - #4800

**Prefer on a per-commit basis, stuff has moved around** Range queries are back!... in the most primitive form possible. No invalidation, no bucketing, no optimization, no nothing. Just putting everything in place. https://github.com/rerun-io/rerun/assets/2910679/a65281e4-9843-4598-9547-ce7e45197995 --- Part of the primary caching series of PR (index search, joins, deserialization): - #4592 - #4593 - #4659 - #4680 - #4681 - #4698 - #4711 - #4712 - #4721 - #4726 - #4773 - #4784 - #4785 - #4793 - #4800

#4785) Title. https://github.com/rerun-io/rerun/assets/2910679/cf2c2748-a461-49fe-8124-c2a94164c956 --- Part of the primary caching series of PR (index search, joins, deserialization): - #4592 - #4593 - #4659 - #4680 - #4681 - #4698 - #4711 - #4712 - #4721 - #4726 - #4773 - #4784 - #4785 - #4793 - #4800

The most obvious and most important performance optimization when doing cached range queries: only upsert data at the edges of the bucket / ring-buffer. This works because our buckets (well, singular, at the moment) are always dense. - #4793 ![image](https://github.com/rerun-io/rerun/assets/2910679/7246827c-4977-4b3f-9ef9-f8e96b8a9bea) - #4800: ![image](https://github.com/rerun-io/rerun/assets/2910679/ab78643b-a98b-4568-b510-2b8827467095) --- Part of the primary caching series of PR (index search, joins, deserialization): - #4592 - #4593 - #4659 - #4680 - #4681 - #4698 - #4711 - #4712 - #4721 - #4726 - #4773 - #4784 - #4785 - #4793 - #4800

…ow-level range queries (#4793)" This reverts commit dc8cf2d.

Range queries used to A) return the frame a T-1, B) accumulate state starting at T-1 and then C) yield frames starting at T. A) was a huge issue for many reasons, which #4793 took care of by eliminating both A) and B). But we need B) for range queries to be context-free, i.e. to be guaranteed that `Range(5, 10)` and `Range(4, 10)` will return the exact same data for frame `5`. This is crucial for multi-tenant settings where those 2 example queries would share the same cache. It also is the nicer-nicer version of the range semantics that we wanted anyway, I just didn't realize back then that it would require so little changes, or I would've gone straight for that. --- Part of the primary caching series of PR (index search, joins, deserialization): - #4592 - #4593 - #4659 - #4680 - #4681 - #4698 - #4711 - #4712 - #4721 - #4726 - #4773 - #4784 - #4785 - #4793 - #4800 - #4851 - #4852 - #4853 - #4856

Simply add a timeless path for the range cache, and actually only iterate over the range the user asked for (we were still blindly iterating over everything until now). Also some very minimal clean up related to #4832, but we have a long way to go... - #4832 --- - Fixes #4821 --- Part of the primary caching series of PR (index search, joins, deserialization): - #4592 - #4593 - #4659 - #4680 - #4681 - #4698 - #4711 - #4712 - #4721 - #4726 - #4773 - #4784 - #4785 - #4793 - #4800 - #4851 - #4852 - #4853 - #4856

Implement range invalidation and do a quality pass over all the size tracking stuff in the cache. **Range caching is now enabled by default!** - Fixes #4809 - Fixes #374 --- Part of the primary caching series of PR (index search, joins, deserialization): - #4592 - #4593 - #4659 - #4680 - #4681 - #4698 - #4711 - #4712 - #4721 - #4726 - #4773 - #4784 - #4785 - #4793 - #4800 - #4851 - #4852 - #4853 - #4856

- Quick sanity pass over all the intermediary locks and refcounts to make sure we don't hold anything for longer than we need. - Get rid of all static globals and let the caches live with their associated stores in `EntityDb`. - `CacheKey` no longer requires a `StoreId`. --- - Fixes #4815 --- Part of the primary caching series of PR (index search, joins, deserialization): - #4592 - #4593 - #4659 - #4680 - #4681 - #4698 - #4711 - #4712 - #4721 - #4726 - #4773 - #4784 - #4785 - #4793 - #4800 - #4851 - #4852 - #4853 - #4856

teh-cmc added 🔍 re_query affects re_query itself do-not-merge Do not merge this PR include in changelog 🔩 data model labels Jan 12, 2024

teh-cmc commented Jan 12, 2024

View reviewed changes

crates/re_viewer/src/ui/visible_history.rs Outdated

Copy link

Member Author

teh-cmc Jan 12, 2024

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

cc @abey79

Wumpf self-requested a review January 15, 2024 10:13

Wumpf approved these changes Jan 15, 2024

View reviewed changes

crates/re_query/tests/archetype_range_tests.rs Show resolved Hide resolved

crates/re_data_store/tests/data_store.rs Outdated Show resolved Hide resolved

teh-cmc force-pushed the cmc/primcache_13_range_stats branch from 77a0113 to a79d00d Compare January 15, 2024 14:51

Base automatically changed from cmc/primcache_13_range_stats to main January 15, 2024 15:17

teh-cmc added 6 commits January 15, 2024 16:19

do not bake latest-at semantics in low-level range APIs

814a7ff

do not bake latest-at semantics in uncache query APIs

c2bc543

cached range test suite shall now pass

f165e7e

update visual history UI accordingly

16093b9

always display single entry plots as a point

204689f

review

07bdde9

teh-cmc force-pushed the cmc/primcache_14_correctness branch from 3dc9114 to 07bdde9 Compare January 15, 2024 15:30

teh-cmc removed the do-not-merge Do not merge this PR label Jan 15, 2024

teh-cmc merged commit dc8cf2d into main Jan 15, 2024
23 of 32 checks passed

teh-cmc deleted the cmc/primcache_14_correctness branch January 15, 2024 15:32

teh-cmc added a commit that referenced this pull request Jan 16, 2024

Revert "Primary caching 14: don't bake LatestAt(T-1) results into l…

6295c43

…ow-level range queries (#4793)" This reverts commit dc8cf2d.

teh-cmc added a commit that referenced this pull request Jan 17, 2024

Revert "Primary caching 14: don't bake LatestAt(T-1) results into l…

f87dec7

…ow-level range queries (#4793)" This reverts commit dc8cf2d.

teh-cmc mentioned this pull request Jan 18, 2024

Primary caching 16: context-free range semantics #4851

Merged

4 tasks

teh-cmc added a commit that referenced this pull request Jan 18, 2024

Revert "Primary caching 14: don't bake LatestAt(T-1) results into l…

754300c

…ow-level range queries (#4793)" This reverts commit dc8cf2d.

This was referenced Jan 18, 2024

Primary caching 17: timeless range #4852

Merged

Primary caching 18: range invalidation (ENABLED BY DEFAULT 🎊) #4853

Merged

Primary caching 19 (final): de-staticify cache globals #4856

Merged

abey79 added the 🚀 performance Optimization, memory use, etc label Feb 7, 2024

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Primary caching 14: don't bake `LatestAt(T-1)` results into low-level range queries #4793

Primary caching 14: don't bake `LatestAt(T-1)` results into low-level range queries #4793

teh-cmc commented Jan 12, 2024 •

edited by github-actions bot

Loading

teh-cmc Jan 12, 2024

Wumpf left a comment

Primary caching 14: don't bake LatestAt(T-1) results into low-level range queries #4793

Primary caching 14: don't bake LatestAt(T-1) results into low-level range queries #4793

Conversation

teh-cmc commented Jan 12, 2024 • edited by github-actions bot Loading

Checklist

teh-cmc Jan 12, 2024

Choose a reason for hiding this comment

Wumpf left a comment

Choose a reason for hiding this comment

Primary caching 14: don't bake `LatestAt(T-1)` results into low-level range queries #4793

Primary caching 14: don't bake `LatestAt(T-1)` results into low-level range queries #4793

teh-cmc commented Jan 12, 2024 •

edited by github-actions bot

Loading