BLOG

Field notes from production AI

Governance, delivery, and engineering. Written by the team that ships agents into regulated enterprises, not by a content calendar.

ENGINEERING

Picking a model per task, with evals to prove it

Model-agnostic is easy to say and hard to operate. Here is how we match models to tasks inside enterprise agents, and the small eval harness that keeps the choice honest.

Read article · 2 min →
DELIVERY

One agent, one job: how we scope automation that ships

Broad AI assistants demo well and deploy badly. The agents that make it to production own one process end to end. Here is the scoping method we use with every customer.

Read article · 2 min →
GOVERNANCE

Why enterprise AI pilots die in security review

Most enterprise AI pilots don't fail on accuracy. They fail on audit trails, access control, and data residency. Here is what security teams actually check, and how to pass.

Read article · 2 min →