Operational bounded quant pipeline; CPU reference measured, GPU comparison requires a kernel receipt
Three orthogonal risk signals per bar — PCA-Risk · TDA-Fracture · HJB-Kelly — with receipt state SIGNED_VERIFIED, honestly labeled SAMPLE_SIGNAL · NOT_LIVE · NO_BACKTEST_VALIDATED. Not a trading instruction.
Ledoit-Wolf shrinkage intensity ρ and the MP noise edge λ⁺ are computed exactly per the cited formulas on a SAMPLE synthetic universe. cuML/cuPy GPU acceleration is ROADMAP (CPU fallback ran here). NO backtest.
Layer 2 · TDA Fracture (β0/β1)
fracture f_t0.0
z-score-1.8717
anomaly |z|>2.5False
β0 (components)1 → 1
β1 (loops)23 → 23
β0 (components) is exact via union-find; β1 is the 1-skeleton cycle rank (E−V+C), an honest fast proxy for the genuine Vietoris-Rips H1 that the GPU giotto-tda/Ripser++ path (ROADMAP) computes exactly. Synthetic SAMPLE windows; z-score baseline is illustrative. NOT calibrated on real data.
Layer 3 · HJB-Kelly Sizing
σ²_eff inflation1.0
Kelly gross exposure60.562005
de-risk ratio1.0
γ, κ0.5, 1.0 (uncalibrated)
Weights auto-compress when fracture/anomaly fire (derisk_ratio<1) — the elegant property of the TDA-Kelly channel. BUT γ and κ are free, UNcalibrated hyperparameters: this is MODELED architecture, NOT a backtested strategy. NEVER a live-trading instruction.
keeps the main GPU from stalling on inline review/draft — best fit for our agent+Auto-Review arch
cloud · NVIDIA NIM (Nemotron 3 Ultra) — frontier/hard tier
wherecloud
sovereignFalse
config
Route via build.nvidia.com NIM through our LiteLLM/RouteLLM gateway
Ultra (550B-A55B) needs ~768GB VRAM (4×GB200-class) — CANNOT run on 2 local GPUs. NEVER claim local Ultra. Verify NVIDIA claims on OUR τ-bench+J/token harness.
Verify the Claims — vendor statement vs bounded local receipt
Claim
NVIDIA published
SZL comparison
Separate local evidence
Nemotron speedup vs prior frontier
up to 5x
— HISTORICAL
0.5217x local Ollama wall-speed ratio vs qwen2.5:3b
Reasoning/accuracy uplift
+30%
16.667 percentage points vs local baseline HISTORICAL
—
Benchmark accuracy
91%
100.0% (6/6 exact match) HISTORICAL
—
Long-context retrieval
1M-token retrieval
1/2 needles; max 2050 evaluated prompt tokens HISTORICAL
—
cuML PCA speedup (quant Layer 1)
10-50x (S&P 500 scale); about 100x genomic
— HISTORICAL
CPU reference p50 11.09 ms; GPU comparison UNAVAILABLE
Ripser++ persistence (quant Layer 2)
up to 30x vs CPU Ripser
— HISTORICAL
CPU stress reference p50 10.378 ms; GPU comparison UNAVAILABLE
The NVIDIA column is a cited vendor claim, not an endorsement. The SZL column comes from a bounded local execution receipt and is not presented as a vendor-scale replication. Signature state is derived from cryptographic verification; unsigned content-addressed evidence is never displayed as signed.