Fuzit
VERIFIED BENCHMARK EVIDENCE — RELEASE v0.0.9
Release Commit
462bf37(@fuzit/cli@0.0.9)

Measured repository intelligence, not benchmark theater.

Fuzit publishes deterministic retrieval metrics plus real-repository timing and selectivity measurements, with reproducible methodology and hardware environment attached.

AMD Ryzen 5 6600H with Radeon Graphics
Node v24.19.0
5/5 Release Audits Passed
FIXED WORKLOAD ACCURACY

Retrieval Quality Metrics

Deterministic rank quality across offline V1 retrieval fixture suites. No LLMs or external APIs involved.

PRECISION
1.0
Ranked evidence precision
RECALL
1.0
Target path coverage
nDCG
1.0
Normalized discounted cumulative gain
MRR
1.0
Mean reciprocal rank
Caption: Deterministic offline V1 retrieval fixture suite (4 test cases, 0 regressions). No LLM or external API involved.
Methodology note: These metrics compare ranked context selection against graded expected paths and do not measure downstream agent task success.
REAL-WORLD CODEBASE PERFORMANCE

Real-Repository Evidence Table

Measured scan latencies, context selection latencies, and file-set reduction across 5 open-source repositories under a fixed 8,000 token budget.

Fixed Token Budget: 8,000 Tokens
RepositoryLanguageCommitDiscoveredSelectedFile-Set ReductionScan MedianContext MedianContext p95Budget Used
fuzitTypeScripte50f099e854499.53%1879.3 ms3560.9 ms3660.0 ms7999/8000
fastifyJavaScript2e81c38b390598.72%1415.7 ms2327.9 ms2544.5 ms7995/8000
flaskPython6a2f545b235996.17%1112.2 ms1804.2 ms1894.1 ms8000/8000
cobraGoadbc881366395.45%726.7 ms973.7 ms1073.1 ms7998/8000
spring-petclinicJava88e37c15131496.95%1012.5 ms1408.3 ms1428.3 ms7996/8000
* Note: fileSetReductionPercent measures (1 - selectedFiles / discoveredFiles). It is strictly a file selectivity metric and is separate from token compression.
CLI STARTUP LATENCY

Binary & Process Overhead

Cold vs repeated CLI process startup measurements and peak memory allocation.

FIRST RUN
616.8 ms
REPEAT MEDIAN
600.6 ms
REPEAT P95
619.6 ms
PEAK WORKING SET
118.4 MB
SYNTHETIC INDEX SCALE

50,000 synthetic file records

Explicit benchmark measuring canonical index reconciliation behavior on a synthetic 50k record set.

Workload TypeSynthetic index reconciliation workload
Index EquivalenceCanonical cold / warm equivalence verified
Reconciliation ModesCold, Warm, 1-file incremental
* Safe label reminder: This is a 50,000 synthetic file record index reconciliation workload, not a real repository scan.
RELEASE GATE VERIFICATION

Release Validation Audit Ledger

All 5 release validation check suites passed prior to publishing version 0.0.9 benchmark evidence.

Full test suite
36.17s
PASS
Security suite
6.36s
PASS
Adversarial suite
4.18s
PASS
Secret audit
15.20s
PASS
Privacy audit
10.21s
PASS
REPRODUCIBILITY & HARDWARE

Methodology & Environment Disclosure

Without complete disclosure of hardware environment, runtime version, and fixed budget constraints, numbers cannot be reproduced.

Release & Software
Version: v0.0.9
Commit: 462bf37
CLI Package: @fuzit/cli@0.0.9
Node Runtime: v24.19.0
Hardware & OS
OS: Windows 11 Home
CPU: AMD Ryzen 5 6600H with Radeon Graphics
Threads: 12 logical cores
Memory: 15.3 GB RAM
Methodology Rules
Token Budget: Fixed 8,000 Tokens
First vs Repeat: Tracked & reported separately
Runs per repo: 7 runs (median & p95 reported)
Baseline mode: --no-index context latency baseline
Fuzit benchmark evidence generated at 2026-08-10T16:01:32.576Z