SZL-Nemo is our sovereign, governed, self-improving agent model — delivered here as a
live skeleton + architecture. Built ON an open base (default Qwen3-32B, Apache-2.0);
governed & sovereign. The differentiator is an auditable MoE: a query is routed to
domain-expert heads by a Λ-governed router (Conjecture 1, advisory floor < 1.0) that reuses
the active-flux router crossover and the RouteLLM Thompson posteriors — and every expert selection emits a
signed DSSE receipt.
Honest by design.
We did NOT train a foundation model from scratch — there is no 550B SZL model and
no local Nemotron-Ultra (cloud tier only). OUR contribution is the governance layer:
governed-MoE domain-expert routing, MTP / speculative-decode default, Reflexion + Voyager + τ-bench
self-improvement, tiered sovereign-local (2-GPU) / cloud-NIM serving, and tamper-evident signed receipts.
sovereign:true only with a live per-GPU gpu_reachable probe; the cloud tier is
always sovereign:false. Every number is MEASURED (live) or ROADMAP — never fabricated.
Λ = Conjecture 1 (advisory, never a pass/fail oracle); trust < 100%; locked-8 @ c7c0ba17; 0 runtime JS CDN;
effectors SIMULATED human-on-loop. Fonts via Google Fonts; no runtime logic/3D CDN.
Governed-MoE Domain-Expert Router the differentiator · signed every selection
"Experts" are domain heads (counter-uas / maritime / governance / code / finance) — an auditable MoE,
not learned FFN experts. Routing fuses the Λ governance score (Conjecture 1) with the active-flux serving
crossover (small/local ⇄ large/cloud) and the RouteLLM Thompson posteriors. Try a query:
Λ —
SIGNED DSSE RECEIPT (route decision)
—
verify—
MTP / Speculative Decoding inference default
Speedup S = (k+1) / (k(1−α)+1) (Leviathan et al. 2022). Draft model serves k tokens,
target accepts at rate α. Reuses Dev C's draft-model wiring; box config is ROADMAP→Forge.
Runs the REAL τ-bench (Dev B) for a deliberately weaker baseline, adopts rule-following
(Reflexion), admits a Voyager skill, re-runs, and signs the measured delta. Score history lives inside receipts.
—
SIGNED DELTA RECEIPT
— (run an iteration)
Serving Tiers honest where / sovereign labels
sovereign:true ONLY with a live per-GPU gpu_reachable probe (Dev C). The cloud NIM
frontier tier (Nemotron Ultra) is always sovereign:false — it needs ~768GB VRAM and cannot run on the 2 GPUs.
—
τ-bench — MEASURED-by-SZL Dev B real suite
SZL τ-bench-style tool-rule-following suite with negative controls; an always-pass agent scores 0,
proving non-triviality. NOT the upstream leaderboard.