645 Commits
Author SHA1 Message Date
andrew 3a82d90914 Work around arm64 clang codegen issues breaking CI
Disable preserve_none under ASan on aarch64: every clang tested (20,
21, trunk 22) miscompiles the continuation chains with that
combination, crashing ~97% of fuzz corpus tests. The chains still work
with the default calling convention. See #38.

Mark the masked Node3 scan results defined for valgrind on aarch64:
clang 21+ lowers the in-bounds tests through flags+csel, which
memcheck models imprecisely, tainting bits the mask provably clears.
See #39.

Skip valgrind tests in the arm64 release job, since -DNVALGRIND
compiles out the client requests the workaround relies on.

Both issues track filing upstream bugs.
2026-06-12 16:35:57 -04:00
andrew 2642d453dc Remove unnecessary branch for interleaved range writes 2024-12-11 21:53:44 -08:00
andrew 7166811387 Fix condition checking for erasing the root 2024-11-21 17:49:54 -08:00
andrew 81323972aa Remove "memcpy common section to next node"
Do it in a simpler, more robust/type safe way
2024-11-21 16:47:05 -08:00
andrew 8694ba8b6a Remove unused field 2024-11-21 16:35:10 -08:00
andrew 0cea5565b5 Add artificial root parent and stop plumbing impl everywhere 2024-11-21 16:32:13 -08:00
andrew 972f16ed8f Remove Node::endOfRange
This should simplify storing leaves directly
2024-11-21 11:55:30 -08:00
andrew 2412684316 childMaxVersion memory needs to be defined
Also remove some dead code, and only memset children to 0 for Node256.
Other nodes detect absence of children other ways.
2024-11-21 10:35:05 -08:00
andrew 8251631087 Fix some unintentional generic usages of getChildAndMaxVersion 2024-11-20 21:50:48 -08:00
andrew 90fb2a9542 Remove some usages of generic getFirstChild 2024-11-20 21:43:47 -08:00
andrew 7c01f8ba0f Remove some usages of maxVersion 2024-11-20 21:24:59 -08:00
andrew 0df2db7f8a Remove some bad unlikely annotations found by -Wmisexpect
Using a profile from server_bench
2024-11-20 19:06:50 -08:00
andrew e125b599b5 Remove freeList, min/max capacity tracking
The freelist doesn't seem to get a good hit rate. Policies other than
capacity = minCapacity did not improve the rate we were resizing nodes,
but did increase memory usage, so get rid of that too. Add a
nodes_resized_total counter.
2024-11-20 14:45:56 -08:00
andrew 3f4d3b685a More valgrind annotations 2024-11-20 13:36:30 -08:00
andrew 4198b8b090 Some prep for leafs-in-parents 2024-11-20 12:20:11 -08:00
andrew 8757d2387c Call prefetch within TaggedNodePointer::getType
Instead of at every call site
2024-11-20 12:07:32 -08:00
andrew 4a22b95d53 Remove state machine transitions "from" Node0
Those aren't used
2024-11-20 11:46:45 -08:00
andrew 03d6c7e471 Allocate maxCapacity instead of minCapacity 2024-11-19 16:13:57 -08:00
andrew ceecc62a63 Disable freelist for macos
It's faster on my mac m1 without the freelist
2024-11-19 14:28:15 -08:00
andrew e5b9c03e77 Restore change that hurt codegen
I don't understand, but this is necessary for good codegen somehow
2024-11-18 14:29:55 -08:00
andrew ee5a84cd7b Remove dead stores 2024-11-15 17:03:29 -08:00
andrew 77262ee2d3 Fix some grammar in a comment 2024-11-15 16:47:31 -08:00
andrew 2777e016ff Be more consistent about TaggedNodePointer vs Node* 2024-11-15 16:30:05 -08:00
andrew 661ffcd843 Explain purpose for prefetches in getFirstChild 2024-11-15 16:29:42 -08:00
andrew 3a34d3cecb Minor style improvements 2024-11-15 16:29:17 -08:00
andrew 189c73e3bd Reduce Node3 size by 8 2024-11-15 15:53:18 -08:00
andrew f5ec9f726a Remove getFirstChildExists
Once we know the type, for Node3 and higher we know the first child
exists anyway
2024-11-15 13:05:11 -08:00
andrew 552fc11c5d Prefetch second child to improve scan performance 2024-11-15 12:45:01 -08:00
andrew 08958d4109 Remove the reverse step for saving scan search path 2024-11-14 16:27:25 -08:00
andrew b78e817e24 Only set HAS_AVX for x86_64 2024-11-12 18:16:30 -08:00
andrew 665a9313a4 Valgrind annotations for new free list 2024-11-12 18:11:09 -08:00
andrew a92271a205 Build with -Wunused-variable 2024-11-12 17:50:27 -08:00
andrew 0dbfb4deae Detect simd headers in c++ instead of cmake 2024-11-12 17:50:27 -08:00
andrew b37feb58dd Require musttail and preserve_none for interleaved 2024-11-12 17:50:27 -08:00
andrew f85b92f8db Improve next{Physical,Logical} codegen 2024-11-09 13:21:36 -08:00
andrew 3c44614311 Allocate from freelist with min/max capacity constraints 2024-11-08 21:35:13 -08:00
andrew 9c1ac3702e Move index closer to start of Node{3,16}
This should slightly improve cache hit rate
2024-11-08 20:54:00 -08:00
andrew 4494359ca2 Make insert_iterations for interleaved writes more closely match 2024-11-04 14:46:30 -08:00
andrew 4eaad39294 Maintain capacity invariant strictly 2024-11-04 13:43:02 -08:00
andrew d6269c5b7c Use interleaved even if preserve_none is missing 2024-11-01 15:12:59 -07:00
andrew 821179b8de Remove dead code 2024-11-01 13:48:32 -07:00
andrew 681a961289 Dispatch on type pairs for end iter 2024-11-01 13:45:41 -07:00
andrew c73a3da14c Dispatch on type pairs for begin iter 2024-11-01 12:46:36 -07:00
andrew 5153d25cce Dispatch on type pairs for common prefix iter 2024-11-01 12:06:48 -07:00
andrew d2ec4e7fae Fix use of wrong member in endIter 2024-11-01 11:17:26 -07:00
andrew ec1c1cf43f Run gc less frequently
But still run it to account for the version potentially increasing
2024-10-31 15:35:00 -07:00
andrew eaad0c69a7 Store Result* pointer instead of index 2024-10-31 12:59:36 -07:00
andrew 309e6ab816 Dispatch on pair of types 2024-10-30 22:59:00 -07:00
andrew 0cce9df8a8 Use constexpr int instead of sizeof 2024-10-30 15:36:46 -07:00
andrew 0df09743da Don't assign to nextRangeWrite twice 2024-10-30 15:33:40 -07:00