Location: Bellevue, WA, USA
Remote: Yes. Prefer in person
Willing to relocate: YES
Technologies: Python, TypeScript, FastAPI, Next.js, React, PostgreSQL, pgvector, Redis, Apache Kafka, AWS, GCP, PyTorch, LLMs, RAG, Prompt Engineering
Résumé/CV: https://drive.google.com/file/d/1dhWtipC3WauUodrj3K0abG8rZn5...
Email: [email protected]
Links: https://divyanshusharma.com | https://github.com/divyanshusharma1709
I am an AI/Full-stack Engineer with 3+ YoE focused on production AI pipelines, inference optimization, and high-precision RAG. Recently as a Founding Engineer at Cascade Intelligence (a16z speedrun), I re-architected a vector matching pipeline that tripled RFP match precision from 20% to 60%. I also built Elliot AI, a proactive companion utilizing hybrid RAG to guarantee <300ms context retrieval across 6+ months of history, and previously engineered an AI narration system for VR surgery syncing LLM inference with frame-accurate playback under 200ms latency.
My foundational infrastructure background includes scaling an async queue-based email system from 400K to 1.5M+ daily sends and driving $2.1M+ in annualized revenue impact through microservice A/B experimentation at Amazon. I am actively seeking senior or founding AI engineering roles to tackle massive system design challenges and build core agentic frameworks, and I require a STEM-OPT transfer (no sponsorship cost).