Role at a glance
- Salary
- Not Disclosed
- Location
- Bengaluru, KA, India
- Work arrangement
- On-site
- Employment
- Full-time
- Experience
- 4–6 years of professional experience in data engineering, ML engineering, or applied AI roles.
Spotted an issue?
We’ll check it against the original posting.
Role Summary
This role designs and operates enterprise data and AI solutions, including scalable pipelines, data products, machine learning models, and GenAI/LLM-powered applications. The work spans enterprise source systems, cloud platforms, data governance, and MLOps to support production-grade data and AI capabilities.
What You'll Do
- Design, build, and maintain ETL/ELT pipelines ingesting structured and unstructured data from SAP S/4HANA, Oracle, Salesforce, and...
- Develop and optimize data products using medallion architecture on platforms such as Databricks.
- Implement data quality, lineage, and governance controls using Purview, Alation, or Unity Catalog.
- Build, deploy, and monitor classification, regression, forecasting, and anomaly detection models.
- Develop RAG pipelines, embeddings, vector databases, and prompt engineering solutions for Copilot, Azure OpenAI, or Bedrock.
- Operationalize models through MLOps practices and deploy data and AI workloads on Azure, AWS, or GCP.
Generated from the employer's posting. Verify important details before applying.
View full postingQualifications
4–6 years of professional experience in data engineering, ML engineering, or applied AI roles; delivered at least 2–3 production-grade data or AI projects end-to-end; experience with ETL/ELT, enterprise data lakes/lakehouses, Databricks, data quality, lineage and governance, Python, PySpark, scikit-learn, TensorFlow, PyTorch, RAG pipelines, embeddings, vector databases, prompt engineering, MLOps, CI/CD, MLflow, Azure ML, SageMaker, Vertex AI, Azure, AWS, GCP and Kubernetes.
Required
- 4–6 years of professional experience in data engineering, ML engineering, or applied AI roles
- Delivered at least 2–3 production-grade data or AI projects end-to-end
- ETL/ELT pipelines
- Enterprise data lakes/lakehouses
- Databricks
- Data quality, lineage and governance
- Python
- PySpark
Preferred
- Working in global team ecosystem
- Experience with enterprise ERP data (SAP S/4HANA, Oracle)
- Integration patterns (CDC, IDoc, OData, APIs)
- Supply chain, procurement, Logistics analytics use cases
- Copilot Studio
- Agentic AI frameworks
Original job description
Content provided by the employer
Original job description
Content provided by the employer
Company Description
At SANDISK, our vision is to power global innovation and push the boundaries of technology to make what you thought was once impossible, possible.
At our core, SANDISK is a company of problem solvers. People achieve extraordinary things given the right technology. For decades, we’ve been doing just that. Our technology helped people put a man on the moon.
We are a key partner to some of the largest and highest growth organizations in the world. From energizing the most competitive gaming platforms, to enabling systems to make cities safer and cars smarter and more connected, to powering the data centers behind many of the world’s biggest companies and public cloud, SANDISK is fueling a brighter, smarter future.
Binge-watch any shows, use social media or shop online lately? You’ll find SANDISK supporting the storage infrastructure behind many of these platforms. And, that flash memory card that captures and preserves your most precious moments? That’s us, too.
We offer an expansive portfolio of technologies, storage devices and platforms for business and consumers alike.
Today’s exceptional challenges require your unique skills. It’s You & SANDISK. Together, we’re the next BIG thing in data
Job Description
- Design, build, and maintain scalable ETL/ELT pipelines to ingest structured and unstructured data from SAP S/4HANA, Oracle, Salesforce, and third-party APIs into enterprise data lakes/lakehouses.
- Develop and optimize data products (medallion architecture) on platforms like Databricks
- Implement data quality, lineage, and governance controls using tools such as Purview, Alation, or Unity Catalog.
- Build, deploy, and monitor ML models (classification, regression, forecasting, anomaly detection) using Python, PySpark, scikit-learn, TensorFlow, or PyTorch.
- Develop GenAI and LLM-powered solutions — including RAG (Retrieval-Augmented Generation) pipelines, embeddings, vector databases (Pinecone, Azure AI Search), and prompt engineering for Copilot/Azure OpenAI/Bedrock.
- Operationalize models through MLOps practices — CI/CD for ML, model versioning, drift monitoring, and automated retraining (MLflow, Azure ML, SageMaker, Vertex AI)
- Deploy and manage data/AI workloads on Azure, AWS, or GCP, leveraging services like ADF, Databricks, Glue, Lambda, Functions, and Kubernetes.
Qualifications
- 4–6 years of professional experience in data engineering, ML engineering, or applied AI roles.
- Proven track record of delivering at least 2–3 production-grade data or AI projects end-to-end
- Working in global team ecosystem is a plus
Preferred Qualifications:
- Experience with enterprise ERP data (SAP S/4HANA, Oracle) and integration patterns (CDC, IDoc, OData, APIs).
- Exposure to supply chain, procurement, Logistics analytics use cases (spend, demand forecasting, supplier risk).
- Familiarity with Copilot Studio, or Agentic AI frameworks.
Additional Information
Sandisk thrives on the power and potential of diversity. As a global company, we believe the most effective way to embrace the diversity of our customers and communities is to mirror it from within. We believe the fusion of various perspectives results in the best outcomes for our employees, our company, our customers, and the world around us. We are committed to an inclusive environment where every individual can thrive through a sense of belonging, respect and contribution.
Sandisk is committed to offering opportunities to applicants with disabilities and ensuring all candidates can successfully navigate our careers website and our hiring process. Please contact us at [email protected] to advise us of your accommodation request. In your email, please include a description of the specific accommodation you are requesting as well as the job title and requisition number of the position for which you are applying.
About the company
Sandisk
Large Enterprise
SanDisk, a subsidiary of Western Digital Corporation, is a leading global provider of flash storage solutions. Founded in 1988, the company specializes in developing high-performance memory cards, USB flash drives, solid-state drives (SSDs), and enterprise storage solutions for devices ranging from smartphones to data centers. Known for its innovation and quality, SanDisk has been instrumental in advancing storage technology, enabling users to store and access vast amounts of data quickly and efficiently. With a commitment to reliability and cutting-edge design, SanDisk continues to empower individuals and businesses worldwide with its advanced storage products.
SanDisk, a subsidiary of Western Digital Corporation, is a leading global provider of flash storage solutions. Founded in 1988, the company specializes in developing high-performance memory cards, USB flash drives, solid-state drives (SSDs), and enterprise storage solutions for devices ranging from smartphones to data centers. Known for its innovation and quality, SanDisk has been instrumental in advancing storage technology, enabling users to store and access vast amounts of data quickly and efficiently. With a commitment to reliability and cutting-edge design, SanDisk continues to empower individuals and businesses worldwide with its advanced storage products.