What to build to EXCEED the state of the art, and exactly what to look for when we fetch. Accountable: researcher. Every row = a filed frontier rung, momentum-ranked from real corpora (OpenAlex 4MB · Crossref 3.8MB) via nx_swcompare_gapmap. Map: /org · graph: /atlas · benches+radars: /compare.
| Lane | Beyond-SOTA opportunity | What to fetch (look for) | Rung |
|---|---|---|---|
| video | end-to-end NEURAL codec · implicit neural representation · learned entropy/RDO · perceptual/generative | DCVC-FM, hyperprior/learned-entropy, VVC/AV1 tool-set, GAN-perceptual metrics — look for RD curves vs x264/x265 at equal SSIM | F611 |
| recall/search | neural RERANK · RRF hybrid BM25+dense · web-scale index | BEIR suite (13 tasks), cross-encoder rerankers, SPLADE/ColBERT — look for nDCG@10 + latency + which signals fuse | F231/236 |
| atlas/recombine | link-prediction on the dep graph · workflow mining · evolutionary recombination | node2vec/GraphSAGE, van der Aalst process mining, AlphaEvolve/OpenEvolve, CodeScene change-coupling, Structure101/Lattix DSM, ArchUnit fitness-functions | F225a |
| living-docs | GraphRAG over the doc-graph · executable/literate docs · staleness-detect | OpenAlex momentum: literate-docs (12) · GraphRAG (6) · doc-staleness (3) · LLM-authoring (3) — look for retrieval grounding + freshness treatment | F250 |
| pm/ROI | autonomous-agent productivity metrics (a new category) · real-time EVM | Jellyfish, LinearB, Swarmia, DX getdx, Cortex — the DORA report, the SPACE paper (Forsgren), DX Core 4; look for source-of-metric + uncertainty/provenance treatment (none tag it = our exceed) | F740-743 |
| chain-of-evidence | PQ signatures · transparency log · ZK proofs · witness quorum | Sigstore/Rekor, in-toto/SLSA, C2PA 2.x, AWS QLDB — interop-EXPORT mappings only (3rd-party substrate refused by sovereignty doctrine); look for the trust-root + revocation model | F707 |
| model/LLM | train-our-own AT SCALE (today: honest TOY) · no-float training | Megatron/DeepSpeed parallelism, K-quant/GGUF, the scaling-law papers — look for tokens/param + the integer-determinism boundary | F235 |
| gpu | first sovereign submit on the 5080 · C0 GEMM · tensor-core paths | CUDA/ROCm kernel patterns, CUTLASS tiling, the 5080 ISA — look for occupancy + the bit-exact-vs-fast tradeoff | F101 |
| civic | legislative tracking · grounded aggregation · corruption signal | Crossref 3.8MB corpus, court-docket + SEC-EDGAR + USPTO feeds — look for citations verifiable vs fabricated (the zombie-loop firewall, atlas-hygiene F263) | F318 |
Corpora fetched sovereignly over our own TLS (nx_https_get): OpenAlex (4MB, 2024-26 works) + Crossref (3.8MB). Per-domain .q = the fetch-query spec; .axes = the coverage axes; nx_swcompare_gapmap ranks opportunities by 2025/26 momentum (liar-killed, envelope-declared, silent-truncation banned). Live radars: video · livingdocs · civic · search · and every /compare/<domain>/frontier.
Accountable = researcher; Responsible = the per-lane researcher + the lane engineer who ships the proof. Consulted = the domain census (referee/librarian); Informed = pm (triages accepted rungs into the board via /plan). Every opportunity here is UNPROVEN by design — it becomes real only when a gate or a run proves it, then it enters the atlas and the next fetch looks further out.