Databricks logo

Databricks

✓ Verified Sponsor

Senior Software Engineer - Distributed Data Systems

🗓️ Posted 2026-08-18 San Francisco, California Hybrid ICT

Company shared salary

NA

Market rate

NA

Based on similar roles (title + domain + location).

About the Company

At Databricks, we are passionate about enabling data teams to solve the world's toughest problems - from making the next mode of transportation a reality to accelerating the development of medical breakthroughs. We do this by building and running the world's best data and AI infrastructure platform so our customers can use deep data insights to improve their business. Founded by engineers - and customer obsessed - we leap at every opportunity to solve technical challenges, from designing next-gen UI/UX for interfacing with data to scaling our services and infrastructure across millions of virtual machines.

Responsibilities

  • Build the next generation distributed data storage and processing systems that outperform specialized SQL query engines in relational query performance while providing expressiveness and programming abstractions for diverse workloads ranging from ETL to data science.
  • Develop Apache Spark, the de facto open source standard framework for big data.
  • Provide reliable and high performance services and client libraries for storing and accessing humongous amounts of data on cloud storage backends such as AWS S3 and Azure Blob Store.
  • Build Delta Lake, a storage management system that combines the scale and cost-efficiency of data lakes, the performance and reliability of a data warehouse, and the low latency of streaming, with higher level abstractions including ACID transactions and time travel.
  • Build Delta Pipelines to orchestrate and operate tens of thousands of data pipelines, providing a higher level abstraction for expressing data pipelines and enabling customers to deploy, test, and upgrade pipelines while eliminating operational burdens.
  • Build the next generation query optimizer and execution engine that is fast, tuning free, scalable, and robust.

Requirements

  • BS (or higher) in Computer Science, a related technical field, or equivalent practical experience.
  • Comfortable working towards a multi-year vision with incremental deliverables.
  • Motivated by delivering customer value and impact.
  • 5+ years of production level experience in either Java, Scala, or C++.
  • Strong foundation in algorithms and data structures and their real-world use cases.
  • Experience with distributed systems, databases, and big data systems (Apache Spark, Hadoop).