Agents & Retrieval
Multi-agent orchestration, tool-calling & routing — plus RAG, GraphRAG and memory infrastructure owned end-to-end.
2+ years shipping production LLM systems — from eval harnesses and red-teaming to retrieval & memory infrastructure.
SCROLLLegal AI, health assistants, retrieval & memory infrastructure — end to end.
Multi-agent orchestration, tool-calling & routing — plus RAG, GraphRAG and memory infrastructure owned end-to-end.
LoRA / QLoRA & DeepSpeed ZeRO distributed training — embeddings, rerankers and low-resource translation.
Eval harnesses, golden sets & LLM-as-judge scoring — red-teaming, prompt-injection testing, OWASP LLM Top 10.
Parse corpora, OCR & structure into DB/CSV-grounded sources.
LoRA / QLoRA adaptation, DeepSpeed ZeRO runs & dataset curation.
Wire agents, tools & routing into a reliable workflow.
Eval harnesses, red-teaming & regression gates before every deploy.
An evidence-grounded health assistant built with Dr. Mohan Tanirru (ASU). I fine-tuned open-source LLMs with LoRA/QLoRA and used frontier models to judge on clinical-dialogue data, own the evaluation pipeline — groundedness scoring, hallucination detection, regression gates — and red-team it against OWASP LLM Top 10, with automatic clinician-referral escalation on high-risk queries.
Multi-agent orchestration routing each request to the right specialist — 42 agents, config-driven.
Compliance agent with hybrid RAG, grounding guardrails and LLM-as-judge evaluation.
Swap LLM providers inside Claude Code — a published npm package, global install.