Software Engineer, Model Routing & Inference

Cursor · New York, NY
full-time mid Posted 3 months ago
Apply Now Stand out: build a proof-of-work pitch →

Free GitHub-based preview. Direct apply stays one click away.

Get weekly job alerts like this →

Hiring for this role?

AI Market Demand Pack · $29 one-time

Compare this role's skills with the full AI hiring market. Get ranked demand, salary bands, leading companies, public source URLs, and a decision brief.

See the live sample →

About this role

Our mission is to automate coding. The first step in our journey is to build the best tool for professional programmers, using a combination of inventive research, design, and engineering. Our organization is very flat, and our team is small and talent dense. We particularly like people who are truth-seeking, passionate, and creative. We enjoy spirited debate, crazy ideas, and shipping code. ABOUT THE ROLE As a Software Engineer on the Model Routing & Inference team at Cursor, you'll build the inference platform that powers every AI interaction in the product. This team owns the full inference path: making Cursor's AI faster, more reliable, and more cost-effective at a scale few teams in the world get to operate at. Every agent session, every tab completion, and every chat message flows through your stack. EXAMPLE PROJECTS INCLUDE... - Building and evolving our inference gateway, a single abstraction over every provider's API semantics, so model onboarding becomes a config change. - Designing intelligent cross-provider failover so no single provider outage causes user-visible degradation. - Designing routing backpressure and admission control so traffic spikes don't cascade into providers. YOU MAY BE A FIT IF - You have deep experience building high-throughput, low-latency distributed systems, especially in inference serving, traffic routing, or real-time data pipelines. - You're comfortable reasoning about cost/performance tradeoffs at scale (GPU utilization, provider economics, capacity planning). - You have strong software engineering fundamentals and enjoy shipping production systems that handle millions of requests. - You make good calls in the gray area: weighing reliability, cost, latency, and user experience when there isn't a single "right" answer. APPLYING If there appears to be a fit, we'll reach to schedule 2-3 short technicals. After, we'll schedule an onsite in our office, where you'll work on a small project, discuss ideas, and meet the team. #LI-DNI

Similar Jobs

Related searches:

On-site Jobs Mid-Level Jobs On-site Mid-Level Jobs Mid-Level AI InfrastructureMid-Level Data EngineeringMid-Level Backend & Systems AI Jobs in New York AI Infrastructure in New YorkData Engineering in New YorkBackend & Systems in New York distributed-systemsdata-pipelineinference

Get jobs like this delivered weekly

Free AI jobs newsletter. No spam.