Repository navigation
perf: release failed high-density route searches - #2894
ShiboSoftwareDev wants to merge 3 commits into
Conversation
|
The latest updates on your projects. Learn more about Vercel for GitHub.
|
Benchmark This PRRun benchmarks by commenting on this PR: Comment Everything after Use Any PR whose title contains |
|
/benchmark --same-machine |
|
/benchmark --same-machine --dataset 18 |
|
/benchmark --same-machine --dataset 24 |
Same Machine Benchmark ResultsBoth revisions ran sequentially in one Blacksmith job on Dataset:
Outcome changes: 0 improved, 0 regressed. DRC issues are totaled across solved samples. Timing percentiles include all samples, with failed and timed-out samples counted at their configured timeout; negative timing deltas are faster. Historical failures without timeout metadata make timing percentiles unavailable. Base pipeline stage timingsAutoroutingPipelineSolver9_PreloadedTraceGraph Recorded timings: 85/85 samples (0 partial, including failed or timed-out samples).
PR pipeline stage timingsAutoroutingPipelineSolver9_PreloadedTraceGraph Recorded timings: 85/85 samples (0 partial, including failed or timed-out samples).
Workflow: View run |
Same Machine Benchmark ResultsBoth revisions ran sequentially in one Blacksmith job on Dataset:
Outcome changes: 0 improved, 0 regressed. DRC issues are totaled across solved samples. Timing percentiles include all samples, with failed and timed-out samples counted at their configured timeout; negative timing deltas are faster. Historical failures without timeout metadata make timing percentiles unavailable. Base pipeline stage timingsAutoroutingPipelineSolver9_PreloadedTraceGraph Recorded timings: 16/16 samples (2 partial, including failed or timed-out samples).
PR pipeline stage timingsAutoroutingPipelineSolver9_PreloadedTraceGraph Recorded timings: 16/16 samples (2 partial, including failed or timed-out samples).
Workflow: View run |
Same Machine Benchmark ResultsBoth revisions ran sequentially in one Blacksmith job on Dataset:
Outcome changes: 0 improved, 0 regressed. DRC issues are totaled across solved samples. Timing percentiles include all samples, with failed and timed-out samples counted at their configured timeout; negative timing deltas are faster. Historical failures without timeout metadata make timing percentiles unavailable. Base pipeline stage timingsAutoroutingPipelineSolver9_PreloadedTraceGraph Recorded timings: 26/26 samples (11 partial, including failed or timed-out samples).
PR pipeline stage timingsAutoroutingPipelineSolver9_PreloadedTraceGraph Recorded timings: 26/26 samples (11 partial, including failed or timed-out samples).
Workflow: View run |
|
/benchmark --same-machine --dataset 18 |
Same Machine Benchmark ResultsThe paired benchmark ended with cancelled before both reports were produced. Workflow: View run |
|
/benchmark --same-machine --dataset 18 |
Same Machine Benchmark ResultsBoth revisions ran sequentially in one Blacksmith job on Dataset:
Outcome changes: 0 improved, 0 regressed. DRC issues are totaled across solved samples. Timing percentiles include all samples, with failed and timed-out samples counted at their configured timeout; negative timing deltas are faster. Historical failures without timeout metadata make timing percentiles unavailable. Base pipeline stage timingsAutoroutingPipelineSolver9_PreloadedTraceGraph Recorded timings: 16/16 samples (2 partial, including failed or timed-out samples).
PR pipeline stage timingsAutoroutingPipelineSolver9_PreloadedTraceGraph Recorded timings: 16/16 samples (2 partial, including failed or timed-out samples).
Workflow: View run |
What changed
Production routing disables search-debug capture, but a failed intra-node grid search was still retained by its portfolio candidate. Dataset 18 sample 002 accumulates 62 failed candidates before another candidate solves the node.
This preserves the failure message and retains the failed solver only when search-debug capture is enabled.
Results
Seven alternating fresh-process A/B pairs using the real Dataset 18 sample-002 large node with the production configuration:
Every run solved in 97,190 iterations with the same output hash. A forced-GC heap snapshot also showed 126,401 fewer retained objects and 17.0 MiB less collection backing storage.
The repository's same-machine Dataset 18 benchmark also reported Memory P50 -15.2%, P80 -8.8%, and P90 -0.7%, with no outcome, DRC, timeout, via, or trace changes. Per-sample process peaks remained noisy (8 lower and 8 higher), so the repeated large-node A/B result above is the direct proof of the lifecycle improvement.
Validation
bun run buildbunx tsc --noEmit