Repository navigation
Canonicalize chunked nested types through the builder [builders-child-stack] - #8967
robert3005 wants to merge 22 commits into
2 benchmarks regressed
⚠️ Unknown Walltime execution environment detected
Using the Walltime instrument on standard Hosted Runners will lead to inconsistent data.
For the most accurate results, we recommend using CodSpeed Macro Runners: bare-metal machines fine-tuned for performance measurement consistency.
⚠️ Different runtime environments detected
Some benchmarks with significant performance changes were compared across different runtime environments,
which may affect the accuracy of the results.
⚡ 28 improved benchmarks
❌ 2 regressed benchmarks
✅ 2067 untouched benchmarks
🆕 10 new benchmarks
⏩ 503 skipped benchmarks1
Warning
Please fix the performance issues or acknowledge them on CodSpeed.
Performance Changes
| Mode | Benchmark | BASE |
HEAD |
Efficiency | |
|---|---|---|---|---|---|
| ❌ | Simulation | take[value_shape/duplicate99/struct8/nonnull/chunks=16/indices=100] |
169.1 µs | 200 µs | -15.44% |
| ❌ | WallTime | dict_canonicalize_gt_u8_neon[1000000] |
487.9 µs | 561.2 µs | -13.06% |
| ⚡ | Simulation | canonicalize[1024, 32] |
52.5 µs | 36.4 µs | +44.24% |
| ⚡ | Simulation | canonicalize[256, 32] |
52.6 µs | 36.5 µs | +44.17% |
| ⚡ | Simulation | chunked_opt_bool_into_canonical[(10, 100)] |
277.7 µs | 199 µs | +39.56% |
| ⚡ | Simulation | canonicalize[1024, 8] |
38 µs | 28.2 µs | +34.85% |
| ⚡ | Simulation | canonicalize[256, 8] |
37.9 µs | 28.2 µs | +34.73% |
| ⚡ | Simulation | canonicalize[16, 8] |
38 µs | 28.2 µs | +34.66% |
| ⚡ | Simulation | canonicalize[16, 32] |
66.4 µs | 50.5 µs | +31.69% |
| ⚡ | Simulation | canonicalize[256, 2] |
35.3 µs | 26.9 µs | +31.02% |
| ⚡ | Simulation | canonicalize[1024, 2] |
35.2 µs | 26.9 µs | +30.82% |
| ⚡ | Simulation | chunked_varbinview_into_canonical[(10, 100)] |
367.6 µs | 295.3 µs | +24.52% |
| ⚡ | Simulation | chunked_opt_bool_into_canonical[(100, 50)] |
248.3 µs | 202.2 µs | +22.8% |
| ⚡ | Simulation | canonicalize[16, 2] |
44.6 µs | 37.3 µs | +19.77% |
| ⚡ | Simulation | chunked_varbin_into_canonical[(10, 100)] |
428.2 µs | 358 µs | +19.6% |
| ⚡ | Simulation | chunked_varbinview_opt_into_canonical[(10, 100)] |
553.8 µs | 474.7 µs | +16.66% |
| ⚡ | Simulation | chunked_opt_bool_into_canonical[(1000, 10)] |
101 µs | 86.9 µs | +16.22% |
| ⚡ | Simulation | take[value_shape/shuffled/struct8/nonnull/chunks=16/indices=100] |
3.3 ms | 2.8 ms | +15.38% |
| ⚡ | Simulation | take_chunked_fsl_sorted[8, 64] |
265 µs | 233.2 µs | +13.67% |
| ⚡ | Simulation | take_chunked_fsl_sorted[32, 64] |
279 µs | 245.8 µs | +13.54% |
| ... | ... | ... | ... | ... | ... |
ℹ️ Only the first 20 benchmarks are displayed. Go to the app to view all benchmarks.
Tip
Investigate this regression by commenting @codspeedbot fix this regression on this PR, or directly use the CodSpeed MCP with your agent.
Comparing claude/chunked-canonical-via-builder-9ze0t6 (e0929fb) with develop (900c1f2)
Footnotes
-
503 benchmarks were skipped, so the baseline results were used instead. If they were deleted from the codebase, click here and archive them to remove them from the performance reports. ↩