
For teams and founders, not interview prep. I build multi-tenant RAG platforms, multi-agent orchestration, and fine-tuning pipelines in production, and maintain open-source work on retrieval design.
Bring an architecture, a prototype that isn't holding up, or a decision you're split on. Typical ground: chunking and retrieval quality, when reranking earns its cost, embedding and vector store choice, tenant isolation, evaluation that isn't vibes, orchestration frameworks versus rolling your own, and where fine-tuning beats prompting versus where it burns budget.
You'll leave with concrete changes and an honest read on what's over-engineered. Please send materials at least 24 hours ahead.