AI Tech Lead Trianz
Led LLM reasoning distillation with DPO and GRPO optimization using Unsloth and Qwen3, achieving a 78% improvement on GSM8K mathematical reasoning benchmarks through a custom DSPy evaluation framework. Fine-tuned Gemma-4 for enterprise office-document generation using QLoRA and GRPO, reducing inference latency by 55% and cloud costs by 65% across more than 200,000 monthly automation requests.
Architected security, safety, and multi-tenant infrastructure for the Concierto multi-agent platform, including prompt-injection detection, MCP isolation, policy enforcement, and extensible guardrails. Led an enterprise GraphRAG coding assistant and dynamic MCP server platform, improving code retrieval accuracy, lowering developer resolution time, and reducing token and context overhead across 50+ production enterprise deployments.