Victor Ramirez · San Francisco Bay Area · Director, Developer & Platform Experience @ Moody's Analytics
I build production AI systems and the platforms that let hundreds of engineers ship them.
RAG pipelines, LLM evaluation frameworks, and agentic workflows: measured, shipped, and operated in production. UC Berkeley MIDS. I teach the same material hands-on through Techqueria and the AI Builders: LatinX Edition podcast.
- 17 yrs
- production systems at scale
- 3 yrs
- AI architecture & evaluation
- 3
- live demos you can use right now
- 6+
- podcast episodes & 3 conference talks
What I build
production, not prototypesRAG pipelines & retrieval systems
Retrieval-augmented generation connecting LLMs to enterprise knowledge at scale.
LLM evaluation frameworks
Recall@k, Precision@k, MRR, groundedness. Rigorous benchmarking, not ad-hoc testing.
Agentic workflows
Multi-turn orchestration, function calling, and MCP tool-use patterns in production.
Developer platforms
Internal platforms, observability rollouts, and incident coordination that keep engineering teams shipping and systems reliable.
Technical program leadership
Readiness gates, risk & dependency management, and cross-functional governance across Platform Engineering, SRE, Security, and BU teams.
Featured work
github.com/ramirez-ai-labsTrustClaw live claude vercel
Autonomous email-summarization agent: Gmail webhook + Claude, every dependency proxied through JFrog Artifactory, with a supply-chain audit trail across 900+ packages.
AI Operating System live agents eval langraph
Four-domain LangGraph orchestration with Claude tool-loop safety circuit breaker, CI-gated eval harness, and deterministic fallback paths. Production-grade agentic system grounding.
RAG Evaluation Lab rag eval python
Fully offline lab for evaluating RAG systems: synthetic datasets, embeddings, vector search, and the retrieval metrics I teach at Techqueria.
Speaking
2025 – 2026Writing
medium.com/@vhr1975What I learned Building a Perplexity API api llm integration
Building with the Perplexity API: lessons on real-time search integration, response streaming, and practical patterns for production LLM applications.
Meet AI-Vic: A Conversational Version of My Portfolio ai-vic llm product
Building the portfolio AI assistant on this site: system prompt grounding, edge inference with Llama 3.3, conversational UX patterns, and deploying AI without a RAG layer.
Introducing the RAG Evaluation Lab rag llm-eval
A practical guide to measuring retrieval quality in RAG systems, covering Recall@k, precision, and groundedness as the evaluation foundations every production AI system needs.
Demystifying Generative AI: My Journey with Techqueria Workshops teaching genai
Behind the curriculum: how I designed hands-on AI workshops for the LatinX engineering community, from LLM fundamentals to production RAG systems and agentic architectures.
Introducing AI Builders: LatinX Edition community podcast
The story behind the podcast: why I started a show amplifying LatinX voices in AI and what I've learned from conversations with practitioners building at the frontier of generative AI.