CAESAR
An agentic workflow evaluation and reliability platform combining generated tests, trace analysis and explicit safety gates.
Selected projects and advisory engagements. Some are open-source, some are client work described at the level of detail that's appropriate.
Selected products, systems, and advisory work across strategy and engineering.
Evaluation, orchestration and infrastructure for agentic systems.
An agentic workflow evaluation and reliability platform combining generated tests, trace analysis and explicit safety gates.
The meta-harness that orchestrates AI coding agents - deterministic dispatch, scope as a contract, cross-vendor review, enforced by exit codes instead of hope.
Open-source agentic AI service mesh - infrastructure for observability, routing, and governance in multi-agent systems.
An agentic kitchen assistant and macro intelligence system that transforms fridge ingredients into structured recipes and real-time nutritional tracking.
A local-first thinking system - markdown notes, wikilink navigation, and a calm editorial UI that evolves from linked notes into a durable knowledge graph.
Native AI IDE that turns intent into specs, plans, and production code through a controlled Intent -> Spec -> Plan -> Code workflow.
Multi-model deliberation platform that routes hard questions through an AI council, then exposes how the final answer was formed.
Consumer AI product that turns playlist exports into psychological profiles and shareable story cards without building a tracking business.
Advisory and framework design for enterprise AI controls, decision rights, and assurance models in regulated environments.
Reusable architecture patterns for agent orchestration, tool safety boundaries, and production monitoring in enterprise AI.