AI Engineer specializing in Large Language Models (LLMs) and Retrieval-Augmented Generation (RAG). I build end-to-end systems that combine LLMs, vector databases, and semantic search to solve real problems at scale.
With a strong background in infrastructure and software engineering, I take projects from prototype to production using Python, FastAPI, and vector stores, and design workflows and agents that orchestrate models and automate decisions in real applications.