Meta
Software Engineering Intern
ML Inference Platform Efficiency · Menlo Park, CA · June–August 2026
Built a production ML inference reliability platform across tens of thousands of serving jobs, reducing investigations from days or weeks to approximately 30 minutes. Architected a deterministic execution engine for AI-generated studies with schema validation, provenance caching, race-safe concurrency, and asynchronous scheduling. Developed LLM-agent workflows with 91% precision and 100% recall on routing evaluation, productionizing 876 runs and approximately 1.47 million classifications in two weeks.