AI/ML Engineer
2025 — PresentIndependent · Applied AI
Design and ship retrieval, ranking and evaluation infrastructure for LLM-backed product features.
- Cut hallucination rate 41% with a grounded retrieval + citation pipeline
- Built an offline eval harness running 1.2k cases on every merge
- Reduced inference cost per request by 3.4x through caching and routing
- PyTorch
- FastAPI
- pgvector
- Docker