Read the architecture

One book. Eleven parts.
Every register accounted for.

The definitive reference is the BRAD_DEVICES_COMPLETE_SPEC.md compendium — 1,733 lines across the whole ecosystem, written and kept current by me. It consolidates 78 per-subsystem documents and opens today — it is the paper trail everything on this site points to.

BRAD_DEVICES_COMPLETE_SPEC.md — the unified specification ON DISK

The one book. Parts 0–10 + appendices: ecosystem blueprint, CPU, GPU, NPU, memory & storage, OS, compute/software, console & products, company, roadmap — with honest Gen 1 / Gen 2 / Vision tagging and a source-of-truth directory tree.

1,733 lines · refreshed Sept 2026 · download the 41-page branded book ↓
The branded specs

Eight specs, one visual language.

Every PDF now wears the brand — gold-dark rules, the official wordmark, and register-level diagrams drawn in one shared TikZ style. Download the set:

bradgfx_arch.pdf SOLID

The GPU architecture — three pipelines drawn as stage diagrams.

11 pages · download ↓

bradisa_spec.pdf SOLID

The ISA — real bit-field figures for RRR / RI / BR / JMP.

11 pages · download ↓

bradisa_v2_ext.pdf SOLID

The vector ISA v2 — 16-bit compressed layout + 256-bit VSET lanes.

16 pages · download ↓

bradxon_platform.pdf DESIGNED

BradXon Gen 1 platform — address-layout bands, MMIO, ANF fabric governor.

17 pages · download ↓

product_portfolio.pdf SOLID

The full product line — brand-hierarchy tree + software portfolio.

16 pages · download ↓

competitive_analysis.pdf SOLID

Whitespace vs the giants — the software moat line is shipped.

10 pages · download ↓

bradcore_vs_intel.pdf DESIGNED

Core-by-core vs the incumbent — benchmark tables, honest rows.

8 pages · download ↓
BradVector Core pipeline

GPU pipeline. The six stages every BradVector core runs — drawn as numbered stage diagrams.

NCC cluster topology

NPU topology. The 2×4 NCC mesh, HB-AIM, scheduler and fabric — §3.2 of the book.

BradRAM + SPMP memory architecture

Memory. BradRAM + SPMP — the two-tier unified pool, §4.1 of the book.

The chapters

78 reference documents, one coherent plan.

Every subsystem has a deep-dive. A selection:

BradVector ISA SOLID

The BradVector instruction set — the ISA every Torox GPU shares.

Public · github.com/BradDevicesOfficial/BradVector

Torox Architecture SOLID

Torox microarchitecture, die naming and the Torox G1 die registry.

Design dossier — GPU architecture

Compute Model SOLID

Compute model — kernels, warps, scheduling, shared memory.

Design dossier — compute model

bradlib.h · BradVector Platform API SHIPPED

The stable host-facing surface: compile → pack → session → launch → run → inspect → breakpoints. One C header, one implementation, exposed 1:1 to the browser via site/js/bradvector.js.

src/include/brad/bradlib.h · runs in this page's wasm

Platform Strategy SOLID

The software platform & strategy — why the toolchain is the moat.

Design dossier — platform strategy

BradChip Platform DESIGNED

SoC platforms, boards, cooling, memory modules.

Design dossier — platform spec

BradRAM / SPMP DESIGNED

Memory die, unified pool, power/security/timing deep-dives.

Design dossier — memory (7 docs)

BradOS DESIGNED

The operating system across Mobile–Compute editions.

Design dossier — OS

BradAxis Console DESIGNED

The console — a Torox GPU in a box, with the Nexus controller.

Design dossier — console

TIFA-SDK · AI toolchain ref API SHIPPED

The Python-first AI weapon — import tifa: one-line optimize(), AutoNeuroStream partition (BNV/BVX/BFX), MoE expert spooler, exact FP8/FP4/INT4 digit formats. Stdlib-only, runnable today.

src/tifa/ · python3 src/tifa/examples/bench.py

BMS · Brad Mobile Services SHIPPED

The cloud-service developer surface — REST spec plus stdlib-only reference server: BradID auth, BradMail, BradHub, BradPay ledger, BradCloud, BradNotify, BradGeo.

src/bms/ · python3 src/bms/reference_server.py

BradDev Kit · the one-binary toolchain SHIPPED

The reference developer CLI — braddev cvt / run / dbg / gpu info / gpu top / timeline / test built on the real shipped stack (bradc → .bvbc → BVRT → BradTimeline/bradgdb) plus the SPMP/EROE/Fabric drivers. All self-checks pass; saxpy & fib run breathable kernels end-to-end.

src/braddev/ · gcc braddev.c + stack + -lm

BradSense-Render — where the receipts are sharpest

The reconstruction engine for the Torox GPU. The hardware (Neuro-Stream units, semantic fabric channel, frame generation) is designed and unfabbed — so the honest move is to ship the software reference first and let the numbers argue. M0 is done; here is exactly what it proves and what it does not.

M0 — the spatial core SHIPPED

Real pixels, not a stub. bradsense_reconstruct() used to zero-fill its output buffer and book performance stats; it now upscales the frame you supply. Bilinear and Catmull-Rom bicubic at arbitrary ratio, driven by the quality-mode table, plus a luma-adaptive sharpen with an anti-ringing clamp and a flat-region gate so it does not amplify noise. It returns -ENODATA rather than inventing a frame.

src/drivers/bradsense.c · src/tests/test_bradsense_pixels.c · ctest green in CI

The numbers M0 earns MEASURED

2× super-resolution PSNR against ground truth: 39.43 dB bilinear, 39.44 dB bicubic, 39.57 dB bicubic + sharpen — all asserted in the test suite, along with byte-determinism, exact corner mapping, an untouched flat field, a crisped step edge, and a per-pixel anti-ringing bound. The gate caught a real bug: the luma mean was divided by 9 in border neighbourhoods with less coverage, sharpening every frame edge.

published baselines, re-checked on every push

M1–M4 — queued, not shipped THE ROAD AHEAD

M1 temporal core: history buffer, motion-vector reprojection, responsive masks so disocclusions do not ghost. M2 the metrics harness that makes it a repeatable receipt. M3 expressing the temporal core as BradVector kernels — the first proof the ISA can run the pipeline the Neuro-Stream units will eventually accelerate. M4 the trained per-object semantic experts. None of these are claimed as working.

Brad Silicon/GPU/BRADSENSE_RENDERER_PLAN.md · BRADSENSE_RENDER_SPEC.md

Where the moat is, honestly THE ARGUMENT

"We use AI to upscale" is the floor, not a differentiator — FSR 4, DLSS 4 and XeSS 2 are all machine-learned. The claim worth making is the object-aware semantic pipeline: geometry and object identity feeding the reconstruction, not a filter over a finished frame. M0 is the baseline that claim is measured against. The algorithm is FSR-3-derived but reimplemented from published behaviour (AMD's FidelityFX SDK is MIT), never cargo-copied and never branded FSR.

the plan states the origin; the registry keeps the product Design/Sil
How to read the tags. SOLID / SHIPPED = the software or doc exists and runs today. DESIGNED = the silicon is fully specced; production is a fab step. VISION = future-physics (honestly labeled). Nothing on this site is dressed up beyond its tag. BradSense-Render is Design / Sil in the company registry: the product waits on silicon, and only the software reference behind it is a shipped receipt.

The book is open. Read a chapter.