about
A small team of engineers who only do production AI.
We're an AI engineering agency. We design and build production AI infrastructure for startups who need to move fast but can't afford to rebuild everything six months later.
We keep our focus tight: LLM systems engineering, retrieval architectures, and multi-agent workflows.
How we work, and why it's this small.
Most agencies grow by adding people and taking on more work. That model makes the senior engineer you met in the pitch disappear by week three.
We took the other route: three clients at a time, the people who scoped the work are the people who write it, and the engagement ends with a handover rather than a renewal.
We specialize
Retrieval architectures, LLM systems engineering, agent workflows. Nothing else.
We keep it small
Three engagements at a time, so the people who scoped it are the people building it.
We build the real thing
Observable, cost-controlled, maintainable by your team after we're gone.
No surprises
You'll always know what we're building, what it costs, and why we made each call.
Our technical stack
We're not loyal to any of it. We pick per project and tell you why.
Orchestration
LangChain · LangGraph · LlamaIndex · CrewAI
Vector stores
Pinecone · Chroma · pgvector · Weaviate
LLM providers
Anthropic · OpenAI · Mistral · Cohere
Infrastructure
Python · FastAPI · PostgreSQL · Redis · Celery
Observability
LangSmith · Langfuse · OpenTelemetry · Prometheus
Deployment
AWS · GCP · Docker · Kubernetes
Open source
The patterns we reuse across engagements, published so you can read them before hiring us.
production-llm-starter
The scaffolding we start every build from — config, tracing, eval harness wired up.
llm-evaluation-framework
How we turn “it feels better” into a number that fails a build.
ai-infrastructure-patterns
Caching, routing, and cost-control patterns, with the trade-offs written down.
multi-agent-orchestration
State, hand-offs, and recovery for agent systems that run unattended.
Talk to the people who'd do the work.
No account manager, no discovery deck. Thirty minutes with an engineer who has shipped this before.
// currently
one slot open this quarter
▸ remote, overlapping European & US hours
▸ engagements start within 2–3 weeks
after that, the next opening is Q4