Thermo Fisher Scientific

Thermo Fisher Scientific

Posted via Workday

Sr. Data Platform Engineer

Posted Oct 1, 2026

Role at a glance

Job function
AI & Data Data Engineering
Salary
Not Disclosed
Location
Bangalore, India
Work arrangement
Hybrid
Employment
Full-time
Experience
At least 10 years of professional experience building analytics platforms, including at least 5 years of hands-on Databricks experience;
Education
Bachelor’s or master’s degree in computer science, Information Technology, Engineering, or a related discipline.

Spotted an issue?

We’ll check it against the original posting.

Log in to report

Role Summary

AI-generated

The Sr. Data Platform Engineer builds and operates secure, scalable data platforms and delivers data, AI, generative AI, and analytics solutions using Databricks, Apache Spark, and cloud technologies. The role spans platform engineering, data engineering, AI/ML enablement, business intelligence, and operational support.

What You'll Do

  • Configure and manage Databricks workspaces, clusters, policies, runtimes, and development, QA/UAT, and production environments.
  • Implement platform security and governance, and automate infrastructure and deployments using Terraform, source control, and CI/CD.
  • Build, maintain, integrate, and optimize Spark and Delta Lake ETL/ELT pipelines and Lakehouse solutions.
  • Prepare data and develop ML workflows and generative AI solutions, then evaluate and monitor their quality, safety, performance, and...
  • Create analytical datasets, semantic models, KPIs, dashboards, and reporting layers, translating business requirements into analytics...
  • Establish monitoring, logging, alerting, troubleshooting, operational support, cross-team standards, and technical documentation.

Generated from the employer's posting. Verify important details before applying.

View full posting

Qualifications

Requires a Databricks certification and production experience with Databricks, Spark/PySpark, Python, SQL, Delta Lake, Lakehouse architecture, cloud platforms, Terraform or similar infrastructure-as-code tools, CI/CD, MLflow, AI/ML/GenAI delivery, Databricks SQL, and BI tools. Also requires strong troubleshooting, optimization, monitoring, and production-support skills.

Required

  • At least 10 years of professional experience building analytics platforms, including at least 5 years of hands-on Databricks experience.
  • A Databricks certification.
  • Production experience with Databricks, Apache Spark/PySpark, Python, SQL, Delta Lake, Lakehouse architecture, workspaces, clusters,...
  • Experience with Azure, AWS, or Google Cloud, including cloud storage, Unity Catalog, IAM/RBAC, secrets management, networking, and...
  • Experience with Terraform or similar Infrastructure-as-Code tools, CI/CD, and source-control practices.
  • Working experience with MLflow and AI/ML/GenAI delivery, including model lifecycle, LLMs, embeddings, vector search, RAG, prompt...
  • Experience with Databricks SQL and BI tools such as Power BI, Tableau, or Looker, including semantic models and KPIs.
  • Strong troubleshooting, performance optimization, monitoring, and production-support skills.

Preferred

  • Experience designing enterprise-scale Databricks Lakehouse platforms and standardized multi-environment deployments.
  • Experience with Databricks Asset Bundles, Apache Airflow, Azure Data Factory, or similar deployment and orchestration frameworks.
  • Experience with streaming technologies such as Spark Structured Streaming, Kafka, or Event Hubs.
  • Experience with advanced MLOps and GenAI tooling such as LangChain, LlamaIndex, Hugging Face, Azure OpenAI, Amazon Bedrock, fine-tuning,...
  • Knowledge of data governance, metadata management, data-quality frameworks, BI governance, privacy controls, and responsible AI practices.
  • Additional Databricks certifications or relevant cloud certifications, and experience in regulated or large enterprise environments.

Original job description

Content provided by the employer

Work Schedule

Standard (Mon-Fri)

Environmental Conditions

Office

Job Description

Join Thermo Fisher Scientific, the world leader in serving science, as a Staff Engineer, Software to make a meaningful impact. In this role, you will provide technical leadership and architectural guidance while developing innovative software solutions that enable our customers to make the world healthier, cleaner, and safer. Working in a supportive, multi-functional environment, you will design and implement sophisticated solutions across our product portfolio, from cloud platforms to scientific instrumentation. You will support the growth of other engineers, drive adoption of best practices, and help shape the technical direction of critical projects. This position offers the opportunity to work with advanced technologies while contributing to groundbreaking scientific discoveries.


Description

Experience: 10+ years of professional experience building analytics platforms, including 5+ years of hands-on Databricks experience

Role: Senior Data Platform Engineer

Primary Skills: Databricks, Apache Spark/PySpark, Python, SQL, Delta Lake, Unity Catalog, MLflow, GenAI, BI, Cloud, CI/CD, Terraform

Role Overview

We are seeking an experienced Sr. Data Platform Engineer to build and operate secure, scalable data platforms and deliver reliable data, AI, generative AI, and analytics solutions using Databricks, Apache Spark, and modern cloud technologies.

The role combines platform engineering, data engineering, AI/ML enablement, business intelligence, and operational support, working closely with engineering, data science, analytics, security, DevOps, architecture, and business stakeholders.

Key Responsibilities

Outcome 1: Secure, scalable, and governed Databricks platform

  • Configure and manage Databricks workspaces, clusters, policies, runtimes, and separate Development, QA/UAT, and Production environments.
  • Implement platform security and governance using Unity Catalog, IAM/RBAC, service principals, secrets, private connectivity, and enterprise access controls.
  • Automate Databricks infrastructure and deployments using Terraform, source control, and CI/CD practices.

Outcome 2: Reliable and high-performing data products

  • Build and maintain scalable ETL/ELT pipelines and Lakehouse solutions using Apache Spark, PySpark, Databricks workflows/jobs, and Delta Lake.
  • Integrate Databricks with enterprise data sources, cloud storage, databases, APIs, and downstream applications.
  • Optimize Spark workloads, queries, clusters, and resource usage for performance, reliability, scalability, and cost.

Outcome 3: Production-ready AI, machine learning, and generative AI solutions

  • Prepare data and build ML workflows covering feature engineering, model training, evaluation, deployment, and lifecycle management using Databricks and MLflow.
  • Develop generative AI solutions using LLMs, prompt engineering, embeddings, vector search, retrieval-augmented generation, and model serving.
  • Evaluate and monitor AI solutions for accuracy, relevance, safety, bias, latency, cost, privacy, and governance.

Outcome 4: Trusted analytics and decision support

  • Create analytical datasets, semantic models, KPIs, dashboards, and reporting layers using Databricks SQL and tools such as Power BI, Tableau, or Looker.
  • Translate business requirements into accurate, performant, user-friendly dashboards and self-service analytics solutions.
  • Ensure analytics solutions follow applicable data governance, security, quality, and accessibility standards.

Outcome 5: Operable and reusable platform capabilities

  • Establish monitoring, logging, alerting, troubleshooting, and operational support for data, AI, generative AI, and analytics workloads.
  • Collaborate across architecture, security, infrastructure, engineering, data science, analytics, and business teams to define practical Databricks standards and best practices.
  • Maintain concise technical documentation for platform configuration, deployment, data pipelines, AI workflows, dashboards, and operational procedures.

Required Technical Skills

  • At least 10 years of professional experience building analytics platforms, including at least 5 years of hands-on Databricks experience; a Databricks certification is required.
  • Strong production experience with Databricks, Apache Spark/PySpark, Python, SQL, Delta Lake, Lakehouse architecture, workspaces, clusters, jobs/workflows, and scalable ETL/ELT pipelines.
  • Experience with Azure, AWS, or Google Cloud, including cloud storage, Unity Catalog, IAM/RBAC, secrets management, networking, and secure connectivity.
  • Experience with Terraform or similar Infrastructure-as-Code tools, CI/CD, and source-control practices.
  • Working experience with MLflow and AI/ML/GenAI delivery, including model lifecycle, LLMs, embeddings, vector search, RAG, prompt engineering, and model serving.
  • Experience with Databricks SQL and BI tools such as Power BI, Tableau, or Looker, including semantic models and KPIs.
  • Strong troubleshooting, performance optimization, monitoring, and production-support skills.

Preferred Skills

  • Experience designing enterprise-scale Databricks Lakehouse platforms and standardized multi-environment deployments.
  • Experience with Databricks Asset Bundles, Apache Airflow, Azure Data Factory, or similar deployment and orchestration frameworks.
  • Experience with streaming technologies such as Spark Structured Streaming, Kafka, or Event Hubs.
  • Experience with advanced MLOps and GenAI tooling such as LangChain, LlamaIndex, Hugging Face, Azure OpenAI, Amazon Bedrock, fine-tuning, or model evaluation.
  • Knowledge of data governance, metadata management, data-quality frameworks, BI governance, privacy controls, and responsible AI practices.
  • Additional Databricks certifications or relevant cloud certifications, and experience in regulated or large enterprise environments, are a plus.

Behavioral Competencies

  • Strong analytical, troubleshooting, and outcome-oriented problem-solving skills.
  • Strong ownership of reliable, secure, scalable, and maintainable solutions.
  • Ability to collaborate effectively across engineering, data science, analytics, architecture, DevOps, security, and business teams.
  • Ability to translate business needs into practical data, analytics, and AI solutions.
  • Clear verbal and written communication, including concise technical documentation.
  • Strong attention to data governance, security, privacy, responsible AI, and operational discipline.

Education

Bachelor’s or master’s degree in computer science, Information Technology, Engineering, or a related discipline.

 

Thermo Fisher Scientific

About the company

Thermo Fisher Scientific

Large Enterprise

Thermo Fisher Scientific is a global leader in serving science, providing a wide range of analytical instruments, reagents, and software solutions for healthcare, pharmaceuticals, and life sciences research. With a commitment to innovation and quality, the company enables customers to make significant advancements in scientific discovery and healthcare diagnostics. Thermo Fisher operates in more than 50 countries and employs over 80,000 people, fostering a collaborative environment aimed at enhancing global health and safety. Through its broad portfolio of brands and technologies, the company plays a vital role in accelerating life-changing research and improving patient outcomes worldwide.