// RESUME
Download the PDF, or print directly from your browser.
Guadalupe Higuera
AI Engineer · Gen AI Platform & LLM Infrastructure
Phoenix, AZ · Guadalupe.Higuera@protonmail.com
I build production GenAI platform infrastructure for 10,000+ users at Wells Fargo — agent orchestration, MCP integrations, multi-agent tool-calling. On the side I fine-tune LLMs, build RAG engines, and write the evals that prove they work.
Professional Experience
Wells Fargo
Chandler, AZ
Software Engineer — GenAI Platform Infrastructure
July 2023 – Present
- Core engineer from inception on a production agentic-AI platform serving 10,000+ users across 6 business lines, including policy, finance, and investments
- Built orchestration APIs routing thousands of daily chats across Claude, Gemini, and GPT models; allowing for model agnostic agent orchestrations allowing users to switch models based on both performance and costs
- Designed and built the configuration APIs for users to create, modify, and deploy their own AI agents
- Developed an MCP server allowing for agents to search SharePoint, Confluence, GitHub and SQL databases; support for multi agent summarization allowing for agents to search relevant data much quicker, speeding up agent orchestration timing
- Own deployment and availability on a weekly release cadence; built self-healing infrastructure on Kubernetes (OpenShift); sustaining 99.99% availability across an enterprise environment
- Mentored junior engineers and interns on both AI engineering and software engineering best practices
Sandhills Global
Scottsdale, AZ
Software Development Intern
December 2021 – February 2023
- Shipped production .NET services and React front-ends for internal and customer-facing applications
Projects
iRacing AI Race Strategist | GitHub | Demo
- Fine-tuned Llama 3.1 8B with QLoRA (PyTorch, Hugging Face) on 9,269 category-balanced synthetic examples using Claude API
- Improved evaluation scores and eliminated hallucinations, measured by a 12 metric weighted scorer along with a blind LLM-as-judge A/B comparison
- Designed event-driven LLM orchestrations across a variety of event racing scenarios for a real-time voice race engineer with 800 – 1,100 ms telemetry to speech latency
TestPilot — Self-Repairing Agentic QA | GitHub
- Built an agentic QA tool (TypeScript, Playwright, OpenAI) that turns plain-English stories into generated Playwright tests, and auto repairs Playwright tests on safe UI drift
- Designed layered AI guardrails: veto-only vision diagnosis, repair guards that preserve test intent and a human approval PR gate; ingests GitHub/Jira stories as an MCP client
Titanic Historical RAG — Contradiction Detection | GitHub | Live
- Built a RAG system using over 3,400+ pages of testimony (40K chunks, 158 witnesses) using OpenAI embeddings, Pinecone and Q&A boundary chunking to surface contradictions between witnesses
- Implemented LLM-as-judge contradiction detection (Claude Haiku 4.5) with structured JSON output, confidence scoring, and a DynamoDB verdict cache
Skills
- GenAI / LLM
- Agentic AI, LLM Orchestration (Semantic Kernel, LangGraph, LangChain, AWS Bedrock Agents, MCP), RAG Systems, Vector Databases, QLoRA Fine-Tuning (PyTorch, Hugging Face)
- LLM Evaluation & Observability
- Eval frameworks, LLM-as-Judge, blind A/B comparison, hallucination measurement & reduction, AI guardrails & human-in-the-loop review, LLM Observability
- Cloud & Infrastructure
- AWS (Bedrock, Lambda, DynamoDB, API Gateway, S3, IAM), Kubernetes (OpenShift), Docker, Terraform, CI/CD, LLMOps, Azure (AZ-900 Certified)
- Languages & Frameworks
- Python, FastAPI, SQL, C#, ASP.NET Core, JavaScript/TypeScript, React, REST APIs, Microservices
Education
Arizona State University
Tempe, AZ
B.S., Computer Science
August 2020 – May 2023