Talent Apply
Log in
All jobs
D

Senior Software Engineer, Model Serving

databricks
San Francisco, California
On-siteUSD 166,000 - 225,000 / year

About this role

About the Role Databricks is seeking a Senior Software Engineer to shape and build the Model Serving platform. You’ll design and implement core systems that enable high-throughput, low-latency inference across CPU and GPU workloads, and collaborate across platform, product, infrastructure, and research teams to deliver a scalable serving platform.

What You'll Do

  • Design and implement core systems and APIs powering Databricks Model Serving, ensuring scalability, reliability, and operational excellence.
  • Drive architectural decisions and trade-offs to optimize performance, throughput, autoscaling, and operational efficiency for CPU and GPU serving workloads.
  • Contribute directly to key components across the serving infrastructure — from model container builds and deployment workflows to runtime systems like routing, caching, observability, and intelligent autoscaling — ensuring smooth and efficient operations at scale.
  • Collaborate across product, platform, and research teams to translate customer needs into reliable and performant systems.
  • Lead technical initiatives that improve latency, availability, and cost-effectiveness across both customer-facing and foundational serving layers.
  • Establish best practices for code quality, testing, and operational readiness, and mentor other engineers through design reviews and technical guidance.

What We're Looking For

  • 5+ years of experience building and operating large-scale distributed systems.
  • Experience in model serving, inference systems, or related infrastructure (e.g., routing, scheduling, autoscaling, observability).
  • Strong foundation in algorithms, data structures, and system design as applied to large-scale, low-latency serving systems.
  • Proven ability to deliver technically complex, high-impact initiatives that create measurable customer or business value.
  • Experience building architecture for large-scale, performance-sensitive CPU/GPU inference systems.
  • Strong communication skills and ability to collaborate across teams in fast-moving environments.
  • Customer-focused mindset with the ability to align implementation details with product goals.
  • Passion for mentoring, growing engineers, and fostering technical excellence.

Compensation & Benefits

  • Local Pay Range $166,000—$225,000 USD.
  • The total compensation package for this position may also include eligibility for annual performance bonus, equity, and the benefits listed above.
  • For more information regarding which range your location is in visit our page here.

Your next opportunity starts here

Prepare, apply, track, interview and get hired — all from one platform, with AI in your corner.

Download app

Or sponsor Premium for someone who's job hunting →