A person stays in charge

Blog

Notes on agents doing the repeatable work, and on the decisions that stay with you.

4 Sep 2026

An Evaluation Protocol for Agent Output — Without Chasing Benchmarks
A chat demo has one promotion rule: someone liked the screenshot. An agentic-first company needs a rule you can run twice and get the same answer.

llm-opsevaluationgovernanceagentic-first

4 Sep 2026

RAG as Corporate Memory — How Agents Stop Inventing Identity
Large models are fluent. Fluency is not memory. If an agent writes a hex code, a legal notice, or a founder bio from “what sounds right,” you do not have a brand. You have a rumor…

ragllm-opsgovernanceagentic-first

10 Aug 2026

Agentic-first companies — agents do the repeatable work, people keep the company.
“Agentic-first” is easy to misread as “fire the humans.” That is not what we mean. In glossary contexts, you may see agent-first used interchangeably; both refer to the same opera…

orchestrationagentic-firstllm-ops

10 Aug 2026

What Is LLM Ops for Agentic-First Companies?
LLM Ops is the discipline of running large-language-model systems in production: deploying, observing, evaluating, routing, retrieving, and gating changes so they stay safe and us…

llm-opsorchestrationroutingragagentic-first

10 Aug 2026

LLM Routing, Data-Center Placement, and RAG as Institutional Memory
An agent army that always calls the same model for every task is like a company that always ships the same person to every meeting. You can do it. You should not.

routingragllm-ops

10 Aug 2026

Local Orchestration and Custom LLMs for Agent Workforces
If your agent army only exists inside a single vendor chat UI, you do not have an operating model — you have a subscription.

orchestrationllm-opsagentic-first

10 Aug 2026

ML Evaluation Loops for Agent Ops — Beyond Chat Demos
Large language models made agent demos easy. They did not make operations easy.

llm-opsevaluation

1 Aug 2026

Brand Kits as Source of Truth — Not Vibes
If three agents and two humans each “remember” a slightly different logo blue, you do not have a brand. You have a rumor mill with good intentions.

raggovernance

1 Aug 2026

Dry-Run by Default — Why Preview Beats Hope
Most content systems treat “publish” as the happy path and “preview” as an optional checkbox. We inverted that.

llm-opsgovernance

1 Aug 2026

Human-in-the-Loop Publishing Without the Theater
“Human in the loop” has become a label people paste on systems that still auto-send with a delay. We mean something stricter: no outbound side-effect without an explicit human dis…

llm-opsgovernancehitl

18 Jul 2026

Agents Can't Govern AI — People Do
We've been building the tools for MeltingFace Presence Bot, and one question keeps coming up in our internal reviews: if agents are writing code, managing approvals, and drafting…

governanceagentic-first

10 Aug 2026

Why AI Governance Matters — A Developer's Perspective
If you are an engineer who has deployed models and tools to production, you already know that putting software in front of real users changes everything. Models drift. Inputs beco…

governancellm-opsagentic-first