Staff Software Engineer- Foundation Model Inference
Databricks
San Francisco, CaliforniaOn-siteFull-time
About this role
<p>P-1930</p> <p>At Databricks, we are passionate about enabling data and AI teams to solve the world&39;s toughest problems — from making the next mode of transportation a reality to accelerating the development of medical breakthroughs. We do this by building and running the world&39;s best data and AI infrastructure platform so our customers can use deep data insights to improve their business. Founded by engineers — and customer-obsessed — we leap at every opportunity to solve technical challenges, from designing next-gen UI/UX for interfacing with data to scaling our services and infrastructure across millions of virtual machines. And we&39;re only getting started.</p> <p>As part of the AI team, you&39;ll build the platforms and products that power everything from data apps, AI agents, model training, model serving, and Vector Search. You&39;ll be joining a high-agency, high-visibility team operating at the frontier of AI infrastructure — with deep ties to research, product, and real-world enterprise use cases. Databricks Mosaic AI is one of our fastest-growing businesses, helping thousands of our customers democratize AI within their organizations. We&39;re building the products and infrastructure that power the next generation of AI.</p> <p>The Foundation Model Inference team is the backbone of Databricks’ generative AI capabilities. We build the infrastructure that enables our customers to serve, scale, and optimize frontier models with enterprise-grade reliability and performance. Our Foundation Model APIs provide a unified platform that gives customers access to LLMs with the governance, flexibility, and scalability required for enterprise production workloads.</p> <p>We are looking for high-agency engineers who are excited to work on powering model inference at enterprise scale.</p> <h3><strong>The impact you will have:</strong></h3> <ul> <li>Build LLM in