Building in public
Changelog
Dated and honest: what's shipped, what's corrected, and what's next. A fuller engineering log lives in the repo's BUILD-LOG.md.
Data2026-07
Corpus Intelligence: from file scorer to corpus operating tool
- Corpus Readiness Board: all ten guidelines scored corpus-wide with a per-file matrix, plus a Remediation Backlog whose fix and accept-risk actions re-score the board, the gates, and the handoff together.
- Version & Duplicate Resolution: pairwise detections grouped into sets with a deterministic keep-latest recommendation; accepted exclusions flow to blocked sources, Build's re-rank, and Govern findings with no extra wiring.
- Corpus Atlas: the similarity map's axes upgraded to true PCA (shared engine with Build's projector), decision overlays (PII rings, staleness dimming, token sizing), and a house-built 3D view.
- Deep signals: parsability measured on the visitor's own upload (extraction yield, encoding damage, boilerplate share), heuristic language profiling, and deterministic topic groups where the human confirms every label.
- Cleaning-to-quality proof: the same baseline retriever run raw vs cleaned against an authored golden set, in-browser, with the measured accuracy delta and stale-evidence share, plus a downloadable readiness dossier.
Operate2026-07
Operate: the 7th stage, day two observability plus the loop back to Frame
- Day two observability added to the Enterprise AI Program lifecycle: the four signal families (system SLOs, model quality canary, RAG freshness and staleness, agent and cost) on a 12 week time axis.
- The engineered week 7 incident: SLOs stay green while the answers decay, silent drift caught by canary evals, not infra dashboards.
- The loop closes: a retrain, reindex, rollback, or rescope decision routes a typed feedback contract back to Frame, Build, Deploy, Realize, and Govern. A lifecycle line becomes a program loop.
- Two downloadable artifacts (weekly ops review, incident report), the first real artifact engine implementations.
Honesty pass2026-07
Post review corrections
- LIVE ready labs (multiagent, structured output) relabeled to authored/illustrative. No LIVE badge without a wired call path.
- Model catalog freshened: other provider entries shown as generic tiers, not version pinned; Anthropic models current.
- Registry count fixed to the 23 new labs; internal QA audit trail corrected.
- Public changelog shipped (this page).
Launch2026-07
Portfolio v1: 23 interactive labs
- Collection 2 · Agent & Protocol (8): MCP playground, loop/failure inspector, orchestration, structured output, context/memory, cost simulator, protocol selection, human in the loop approval.
- Collection 3 · Business of AI (5): portfolio dashboard, build versus buy, cost forecaster, vendor monitor, ROI builder.
- Collection 4 · Engagement Leadership (10): adoption, stakeholders, capacity, RAID radar, compliance, talent, RFP, estimation, onboarding, exec comms.
- Layer 0 Competency Map landing, driven by a shared labs registry.
Foundations2026-07
Shared spine
- @labs/kit: dated model, pricing, and protocol config plus the registry that auto updates the map as labs ship.
- One design system across every new collection (matches the Command Center).
- Credibility block on every lab: badge · freshness stamp · steering takeaway · how it's built · limitations.
RoadmapNext
In flight
- Genuine LIVE calls on the flagship agent labs (once the deploy host is set).
- Use Case Layer: 3 real world, cross industry scenarios per lab plus an Industry Atlas.
- Collection index pages (the toolkit / gallery / control room structures).
- Shareability and accessibility: per lab OG images, sitemap, full accessibility pass.
Honest by design: every lab states whether it's LIVE or SIMULATED, and every number expands to its formula.