Databricks

Databricks

Posted via Greenhouse

Staff Software Engineer - Ingestion

Posted Aug 5, 2026

Role at a glance

Salary
Not Disclosed
Location
Bengaluru, India
Work arrangement
On-site
Employment
Full-time
Experience
15+ years industry experience building and supporting large-scale distributed systems.

Spotted an issue?

We’ll check it against the original posting.

Log in to report

Role Summary

AI-generated

The Lakeflow Connect team builds point-and-click connectors that ingest data from enterprise applications, databases, cloud storage, message queues, and local files into the Databricks Lakehouse. This role focuses on database internals and efficient extraction from OLTP systems while minimizing production-system load, supporting scalable Data and AI workflows across Databricks surfaces.

What You'll Do

  • Build a highly scalable, available, and fault-tolerant engine processing hundreds of TB of data daily across thousands of customers.
  • Perform low-level systems debugging, performance measurement, and optimization on large production clusters.
  • Design architecture, influence the product roadmap, and take ownership of new projects.
  • Prevent and investigate production issues.
  • Plan and lead complicated technical projects involving several teams within the company.
  • Mentor others, lead sprint planning, delegate work and assignments, and participate in project planning as a Technical Team Lead.

Generated from the employer's posting. Verify important details before applying.

View full posting

Qualifications

15+ years of experience building and supporting large-scale distributed systems; experience with database replication, backup, or transaction recovery at a major database vendor; strong foundation in algorithms and data structures; experience driving company initiatives toward customer satisfaction.

Required

  • 15+ years industry experience building and supporting large-scale distributed systems
  • Experience in database replication, backup, or transaction recovery at one of the major database vendors
  • Strong foundation in algorithms and data structures and their real-world use cases
  • Experience driving company initiatives towards customer satisfaction

Original job description

Content provided by the employer

P-375

At Databricks, we are passionate about enabling data teams to solve the world's toughest problems - from making the next mode of transportation a reality to accelerating the development of medical breakthroughs. We do this by building and running the world's best data and AI infrastructure platform so our customers can use deep data insights to improve their business.

Ingesting data into the Lakehouse is a strategic area of investment for Databricks and a key enabler for Data and AI workflows. Lakeflow Connect is looking to solve this problem by providing ready-to-use, point-and-click connectors for a wide variety of sources, including enterprise applications (like Salesforce, Workday, ServiceNow, SharePoint), databases (e.g., SQL Server), cloud storage, message queues, and local files.

In addition to being an important part of Lakeflow and Data Engineering, Connect is also a key platform capability. Every surface in Databricks (Dashboards, Notebooks, SQL, AI) requires ingestion capabilities and the lead for this role will need to work closely with other products to embed Connect into these surfaces.

We are looking for engineers with experience in core Database internals to join our Lakeflow Connect team. A key part of Connect is to extract data from OLTP systems while imposing minimal load on production systems. To do this efficiently we are building systems that use techniques such as incremental data capture, log parsing, etc. We are looking for engineers who continue to be hands on and are looking to make a large impact on an important problem for the company.

The Impact you will have:

  • Solve real business needs at large scale by applying your software engineering.
  • Deliver a highly scalable, available, and fault-tolerant engine processing hundreds of TB of data daily across thousands of customers
  • Low level systems debugging, performance measurement & optimization on large production clusters.
  • Build architecture design, influence product roadmap, and take ownership and responsibility over new projects
  • Use your deep experience to help prevent and investigate production issues.
  • Plan and lead complicated technical projects that work with several teams within the company.
  • A strong influencer or driver in the organization’s roadmap and direction.
  • Lead a TLG or similar review committee, or initiate and sustain an org/eng-wide initiative driven by engineering needs.
  • Break down complex problems quickly into potential solutions, knowns, and unknowns, and de-risk (through prototyping/validation).
  • Contribute as a Technical Team Lead by mentoring others, lead sprint planning, delegating work and assignments to team members and participate in project planning.

What we look for:

  • 15+ years industry experience building and supporting large-scale distributed systems.
  • Experience in areas like Database replication, backup, transaction recovery at one of the major database vendors (Microsoft SQL Server , Oracle, IBM etc).
  • Comfortable working towards a multi-year vision with incremental deliverables.
  • Motivated by delivering customer value and impact.
  • Strong foundation in algorithms and data structures and their real-world use cases.
  • Experience driving company initiatives towards customer satisfaction.

About Databricks

Databricks is the Data and AI company. More than 20,000 organizations worldwide — including adidas, AT&T, Bayer, Block, Mastercard, Rivian, Unilever, and 70% of the Fortune 500 — rely on the Databricks Data + AI Platform to build and scale data and AI apps, analytics and agents. Headquartered in San Francisco with 30+ offices around the globe, Databricks offers a unified platform that includes Genie, Lakebase, Agent Bricks, Lakeflow, Lakehouse, and Unity Catalog. To learn more, follow Databricks on LinkedIn, X, YouTube, and Instagram.

Benefits

At Databricks, we strive to provide comprehensive benefits and perks that meet the needs of all of our employees. For specific details on the benefits offered in your region click here.

Our Commitment to Diversity and Inclusion

At Databricks, we are committed to fostering a diverse and inclusive culture where everyone can excel. We take great care to ensure that our hiring practices are inclusive and meet equal employment opportunity standards. Individuals looking for employment at Databricks are considered without regard to age, color, disability, ethnicity, family or marital status, gender identity or expression, language, national origin, physical and mental ability, political affiliation, race, religion, sexual orientation, socio-economic status, veteran status, and other protected characteristics.

Compliance

If access to export-controlled technology or source code is required for performance of job duties, it is within Employer's discretion whether to apply for a U.S. government license for such positions, and Employer may decline to proceed with an applicant on this basis alone.

Applicant Privacy Notice

Databricks

About the company

Databricks

Large Enterprise

Databricks is a cloud-based data platform that specializes in providing solutions for data engineering, machine learning, and analytics. Founded in 2013 by the original creators of Apache Spark, the company enables organizations to unify data processing and analysis workflows, facilitating collaboration for data professionals. Databricks offers an integrated environment for data scientists, engineers, and business analysts, streamlining the process of building and deploying AI and data-driven applications. With a strong emphasis on simplifying big data management, Databricks helps businesses harness the power of their data to drive innovation and insights.