๐ Open to Work
Ashwin Sreedhar
ashwin-sreedhar
ยท
AI & ML interests
LLM inference systems and the data around them. I build emberserve, a from-scratch inference engine with paged KV cache, prefix caching, continuous batching and an OpenAI-compatible server, benchmarked against vLLM on A100s and Runpod Serverless. serverless-lakehouse is the PySpark + Delta medallion pipeline over those benchmarks โ cold starts, FlashBoot, worker boot anatomy, cost per request โ with the gold layer served as a Space here. MS Computer Engineering, Purdue (Dec 2026); previously backend at Gallo.
Recent Activity
updated a Space about 15 hours ago
ashwin-sreedhar/serverless-lakehouse updated a dataset 6 days ago
ashwin-sreedhar/runpod-serverless-benchmarks published a dataset 6 days ago
ashwin-sreedhar/runpod-serverless-benchmarksOrganizations
None yet