Staff Software Engineer, Foundation Model Inference
Databricks
San Francisco, CaliforniaOn-siteFull-time
About this role
<p><strong>P-1930</strong></p> <p>At Databricks, we are passionate about enabling data and AI teams to solve the world&39;s toughest problems — from making the next mode of transportation a reality to accelerating the development of medical breakthroughs. We do this by building and running the world&39;s best data and AI infrastructure platform so our customers can use deep data insights to improve their business. Founded by engineers — and customer-obsessed — we leap at every opportunity to solve technical challenges, from designing next-gen UI/UX for interfacing with data to scaling our services and infrastructure across millions of virtual machines. And we&39;re only getting started.</p> <p>As part of the AI team, you&39;ll build the platforms and products that power everything from data apps, AI agents, model training, model serving, and Vector Search. You&39;ll be joining a high-agency, high-visibility team operating at the frontier of AI infrastructure — with deep ties to research, product, and real-world enterprise use cases. Databricks Mosaic AI is one of our fastest-growing businesses, helping thousands of our customers democratize AI within their organizations. We&39;re building the products and infrastructure that power the next generation of AI.</p> <p>We&39;re hiring across multiple teams in our AI Engineering org, including the <strong>FMAPI (Foundation Model APIs)</strong> team — the unified serving layer for large language models across real-time and batch inference, powering model inference at enterprise scale. We are looking to hire high-agency engineers who bridge the gap between technical execution and product strategy.</p> <h3><strong>The impact you will have:</strong></h3> <ul> <li>Build LLM infrastructure powering large-scale inference workloads for customers through partner models (OpenAI, Anthropic, Gemini) and self-hosted model