AI Engineer
Juristo AI Labs
- Built and scaled a production Next.js AI platform spanning APIs, authentication, background jobs and frontend delivery for 80,000+ users.
- Engineered a 4-stage RAG pipeline — document parsing, chunking, embeddings, retrieval — to deliver context-grounded legal responses.
- Integrated LLM inference through OpenRouter behind a multi-agent routing layer that dispatches queries to the right tools and workflows.
- Cut LLM inference cost 40% with prompt caching and selective routing of suitable workflows to lighter models.