PortfolioChangelog

Building in public

Changelog

Dated and honest: what's shipped, what's corrected, and what's next. A fuller engineering log lives in the repo's BUILD-LOG.md.

Data2026-07

Corpus Intelligence: from file scorer to corpus operating tool

  • Corpus Readiness Board: all ten guidelines scored corpus-wide with a per-file matrix, plus a Remediation Backlog whose fix and accept-risk actions re-score the board, the gates, and the handoff together.
  • Version & Duplicate Resolution: pairwise detections grouped into sets with a deterministic keep-latest recommendation; accepted exclusions flow to blocked sources, Build's re-rank, and Govern findings with no extra wiring.
  • Corpus Atlas: the similarity map's axes upgraded to true PCA (shared engine with Build's projector), decision overlays (PII rings, staleness dimming, token sizing), and a house-built 3D view.
  • Deep signals: parsability measured on the visitor's own upload (extraction yield, encoding damage, boilerplate share), heuristic language profiling, and deterministic topic groups where the human confirms every label.
  • Cleaning-to-quality proof: the same baseline retriever run raw vs cleaned against an authored golden set, in-browser, with the measured accuracy delta and stale-evidence share, plus a downloadable readiness dossier.
Operate2026-07

Operate: the 7th stage, day two observability plus the loop back to Frame

  • Day two observability added to the Enterprise AI Program lifecycle: the four signal families (system SLOs, model quality canary, RAG freshness and staleness, agent and cost) on a 12 week time axis.
  • The engineered week 7 incident: SLOs stay green while the answers decay, silent drift caught by canary evals, not infra dashboards.
  • The loop closes: a retrain, reindex, rollback, or rescope decision routes a typed feedback contract back to Frame, Build, Deploy, Realize, and Govern. A lifecycle line becomes a program loop.
  • Two downloadable artifacts (weekly ops review, incident report), the first real artifact engine implementations.
Honesty pass2026-07

Post review corrections

  • LIVE ready labs (multiagent, structured output) relabeled to authored/illustrative. No LIVE badge without a wired call path.
  • Model catalog freshened: other provider entries shown as generic tiers, not version pinned; Anthropic models current.
  • Registry count fixed to the 23 new labs; internal QA audit trail corrected.
  • Public changelog shipped (this page).
Launch2026-07

Portfolio v1: 23 interactive labs

  • Collection 2 · Agent & Protocol (8): MCP playground, loop/failure inspector, orchestration, structured output, context/memory, cost simulator, protocol selection, human in the loop approval.
  • Collection 3 · Business of AI (5): portfolio dashboard, build versus buy, cost forecaster, vendor monitor, ROI builder.
  • Collection 4 · Engagement Leadership (10): adoption, stakeholders, capacity, RAID radar, compliance, talent, RFP, estimation, onboarding, exec comms.
  • Layer 0 Competency Map landing, driven by a shared labs registry.
Foundations2026-07

Shared spine

  • @labs/kit: dated model, pricing, and protocol config plus the registry that auto updates the map as labs ship.
  • One design system across every new collection (matches the Command Center).
  • Credibility block on every lab: badge · freshness stamp · steering takeaway · how it's built · limitations.
RoadmapNext

In flight

  • Genuine LIVE calls on the flagship agent labs (once the deploy host is set).
  • Use Case Layer: 3 real world, cross industry scenarios per lab plus an Industry Atlas.
  • Collection index pages (the toolkit / gallery / control room structures).
  • Shareability and accessibility: per lab OG images, sitemap, full accessibility pass.

Honest by design: every lab states whether it's LIVE or SIMULATED, and every number expands to its formula.