Tribunal: When AI Decisions Need a Ledger
July 6, 2026. Tribunal turns high-stakes AI output into a reviewable decision record: blind proposals, sealed commitments, critique, vetoes, dissent, ratification, and a hash-chained ledger.
Pablo Zavala · AI Safety Evaluation · Research Engineering
Notes on AI governance, evaluation, economics, and institutional design by Pablo Zavala.
July 6, 2026. Tribunal turns high-stakes AI output into a reviewable decision record: blind proposals, sealed commitments, critique, vetoes, dissent, ratification, and a hash-chained ledger.
July 6, 2026. Heard.now frames civic AI as a checkable listening loop: residents speak plainly, privacy gates protect raw input, sample limits accompany each finding, and campaigns show what changed.
July 6, 2026. JROS turns job search into governed resource work: sources, claims, packets, validation, and human approval gates over application-volume automation.
July 6, 2026. LectureForge turns authorized teaching material into private, inspectable course-book drafts while keeping rights, privacy, review, and publication authority explicit.
July 4, 2026. AEGIS frames cross-owner agent work as a contract, verifier commitment, bid record, receipt, dispute path, settlement authority packet, and execution receipt, with launch claims bound to evidence.
July 2, 2026. Safe MarketUniverses uses finance as a compact lab for a practical agent-safety question: can model-emitted uncertainty guide scarce human review toward decisions where review would have helped?
July 2, 2026. Why model confidence can be calibrated on average but still fail the operational question: which agent decisions deserve scarce human review.
March 24, 2026. GLP-1 drugs work while public coverage stalls, and the gap reads best as a behavioral-economics failure inside institutions: fiscal-score anchoring, an inherited status-quo default, and loss-averse delay that a multi-attribute objective and two framing interventions can correct.
March 16, 2026. Teaching AI governance through Colab notebooks where Run all works deterministically on free hardware, with fixed seeds, dated snapshots, sample caps, and provider-agnostic fallbacks that let students audit every claim, and a structured mini-brief that turns each lab into an argument from question to method to limitation to recommendation.
February 15, 2026. A formal finance manuscript makes time a contract, so every decision reads only data published before it, look-ahead bias vanishes by construction, and each policy becomes a deterministic, auditable artifact, with a conceptual bridge to verifiable AI decisions.
February 10, 2026. AI helps when the system sharpens judgment. Polished output becomes a fluency trap when output replaces the verification, writing, reasoning, and cognitive habits that build human capital.
July 10, 2025. A four-page evidence memo distills the empirical record on U.S. gigabit fiber, grading every claim as causal or descriptive: Chattanooga's $2.69 billion municipal payoff, a roughly 20% incumbent speed response to Google Fiber entry, 2% and nearly 9% housing premiums, speed thresholds for business births, recurring adoption gaps, and a deployment playbook BEAD should enforce.
November 11, 2024. A graduate exercise on synthetic pretrial data shows disparate error rates across race that aggregate fairness summaries miss, a failure mode a safety review should catch and an auditable decision record would surface.
October 13, 2024. A design proposal rather than a built system, for a Spanish, Quechua, and Shuar legal and business advisor, framed around low-resource-language safety, hallucination risk, and human review.