Yashwanth Sai
Actively seeking SWE/SDE Roles | AI Engineer | CUDA Programmer | LLM Tuning | vLLMs
- Role
- Ai Engineer at Stealth Startup
- Location
- New York, NY, US
- LinkedIn followers
- 500 followers
Experience
Ai Engineer
May 2025 — Present · Los Angeles, CA, US
Deployed and optimized AI model serving infrastructure supporting multiple model types using vLLM and PyTorch on GPUs, achieving 90% GPU utilization and high throughput inference for scalable customer applications.·Implemented a microservices based AI deployment pipeline using Docker containers and LiteLLM proxy routing, reducing deployment time by 60% and automating model orchestration across 20+ LLMs, improving inference reliability by 30% and enabling cost efficient scaling for customer applications.·Optimized AI model efficiency using LoRA/QLoRA quantization, reducing VRAM usage up to 50% and enabling deployment on cost effective GPUs, saving 30% on infrastructure costs, while fine-tuning models for multimodal tasks·Delivered multimodal AI capabilities including OCR, speech synthesis, and vision processing via SkyPilot’s intelligent orchestration to manage cost efficient infrastructure and enabling seamless integration for diverse customer use cases.
Education
PES College of Engineering, Mandya
Bachelor of Engineering - BE, Computer Science
University of California, Riverside
Masters in Computer engineering
Find verified contacts for anyone on LinkedIn
Unifers gives sales teams verified emails and direct dials, enriched profiles, and outreach that lands in the inbox.
Free plan included · No credit card required
This profile is compiled from publicly available professional sources. Unifers is not affiliated with or endorsed by LinkedIn. Request removal of this profile.