Priya Nair

Staff Engineer · 2 posts

I build the serving layer — batching, caching, and squeezing throughput out of every GPU.

Follow

Ship inference that scales

Deploy your first model on Nexora in minutes. Pay only for what you run.