Amir Azadfar

AI Systems Engineer

I build the layer that makes AI features cheap and safe to ship — LLM runtimes, governed agent platforms, and the retrieval underneath.

$ focus --on llm-runtimes, agents, semantic-search, prod-infra

Most recently: a governed agent runtime taken from an empty repository to two authorized staging deployments in five weeks, solo — 99K lines, 223 routes, 1,453 tests, and a human approving every consequential write. Open to senior and founding AI engineering roles.

01 / Projects

Things I've built

All projects
02 / Research

Published work

All research
03 / Stack

Stack & expertise

LLM / Retrieval / Search

LLM RuntimesAgentic WorkflowsTool CallingStructured Output + RepairRAGHybrid Retrieval (BM25 + Semantic + RRF)QdrantVector EmbeddingsPrompt/Response Caching

Governance & Reliability

Policy EnginesHuman-in-the-Loop GatesBehavior CertificationEval HarnessesAudit TrailsIdempotency & Exactly-Once DeliveryCost/Budget LedgersCircuit BreakersDeterministic Replay

Backend & Systems

PythonFastAPIasyncioKafkaRedisDockerNginxREST APIsSSE StreamingMicroservices

AI / Machine Learning

PyTorchVision TransformersGraph Neural NetworksScikit-learnStable-Baselines3HuggingFace TransformersSentenceTransformersspaCyWeights & Biases
04 / Writing

Recent posts

All posts