Mira Chen

Infrastructure Lead · 2 posts

I work on inference scheduling and GPU efficiency at Nexora. Before this I spent eight years on distributed systems. I write about the unglamorous engineering that makes models fast in production.

Follow

Ship inference that scales

Deploy your first model on Nexora in minutes. Pay only for what you run.