Perplexity

Perplexity

Posted via Ashby

Member of Technical Staff (Software Engineer, Data Platform)

Posted Aug 5, 2026

Role at a glance

Salary
$220K – $405K/yr
Location
San Francisco, United States Palo Alto, United States
Work arrangement
On-site
Employment
Full-time
Experience
5+ years (Senior) or 8+ years (Staff) of software engineering experience.

Spotted an issue?

We’ll check it against the original posting.

Log in to report

Role Summary

AI-generated

The Data Platform team owns Perplexity’s end-to-end data lifecycle, including ingestion, processing, storage, and serving for product features, analytics, experimentation, AI workloads, and the company’s data lake. In this senior/staff role, the engineer will shape architecture, set standards, and drive the long-term technical direction of the data ecosystem.

What You'll Do

  • Design and operate large-scale batch and streaming data pipelines for product features, AI training and evaluation workflows, analytics,...
  • Build event-driven and streaming systems for real-time ingestion, transformation, and delivery, alongside batch frameworks for...
  • Lead the architecture of data orchestration, including scheduling, dependency management, retries, SLAs, and end-to-end observability.
  • Set and enforce guarantees for data correctness, freshness, lineage, and recoverability across evolving data systems.
  • Build self-serve data platforms that enable engineers, data scientists, and analysts to discover data, define contracts, and create and...
  • Drive architectural decisions across storage, compute, orchestration, and data APIs while mentoring engineers and reviewing designs.

Generated from the employer's posting. Verify important details before applying.

View full posting

Qualifications

Strong experience building production data infrastructure systems; hands-on experience with batch and/or streaming data processing at scale; deep familiarity with Airflow, Dagster, or similar; proficiency in Python and at least one additional backend language; experience supporting ML/AI workflows, training pipelines, or evaluation systems; familiarity with data quality, lineage, observability, and governance tooling; prior ownership of internal platforms used by many teams.

Required

  • Strong experience building production data infrastructure systems.
  • Hands-on experience with batch and/or streaming data processing at scale.
  • Deep familiarity with data orchestration systems (Airflow, Dagster, or similar).
  • Proficiency in Python and at least one additional backend language (Go, TypeScript, etc.).
  • Strong systems thinking around reliability, latency, cost, and complexity tradeoffs.
  • Experience supporting ML/AI workflows, training pipelines, or evaluation systems.
  • Familiarity with data quality, lineage, observability, and governance tooling.
  • Prior ownership of internal platforms used by many teams.

Original job description

Content provided by the employer

About the Role

The Data Platform team owns the end-to-end data lifecycle at Perplexity, from ingestion through processing, storage, and serving, powering product features, analytics, experimentation, AI workloads, and the company’s data lake.

The team defines the architecture for batch and streaming systems, the orchestration and observability stack, and a self-serve data platform, while thoughtfully combining platforms such as Databricks and Snowflake with open-source technologies including Spark, Kafka, Flink, Airflow, Dagster, dbt, Iceberg, Delta Lake, and ClickHouse.

In this senior/staff role, you will shape architecture, set standards, and drive the long-term technical direction of Perplexity’s data ecosystem.

Key Responsibilities

  • Design and operate large-scale batch and streaming data pipelines that directly power Perplexity product features, AI training and evaluation workflows, analytics, and experimentation.

  • Build event-driven and streaming systems (Kafka, Kinesis, PubSub, or similar) for real-time ingestion, transformation, and delivery, alongside batch frameworks for backfills, aggregations, and offline computation.

  • Lead the architecture of data orchestration using tools like Airflow or Dagster, owning scheduling, dependency management, retries, SLAs, and end-to-end observability for critical data flows.

  • Set and enforce guarantees for data correctness, freshness, lineage, and recoverability, designing systems that handle rapid scale growth, partial failures, and evolving schemas without disrupting AI workloads or product experiences.

  • Build self-serve data platforms that let engineers, data scientists, and analysts safely discover data, define contracts, and create and operate their own pipelines with minimal friction.

  • Improve developer experience through better abstractions, opinionated paved paths, and standards for data modeling, testing, validation, and deployment, treating the data platform as a product used by many teams.

  • Drive architectural decisions across storage, compute, orchestration, and data APIs, partnering closely with product engineering and data science to align the data ecosystem with Perplexity’s roadmap.

  • Mentor engineers, review designs, and raise the technical bar for data infrastructure through thoughtful feedback, documentation, and hands-on collaboration.

Qualifications

  • 5+ years (Senior) or 8+ years (Staff) of software engineering experience.

  • Strong experience building production data infrastructure systems.

  • Hands-on experience with batch and/or streaming data processing at scale.

  • Deep familiarity with data orchestration systems (Airflow, Dagster, or similar).

  • Proficiency in Python and at least one additional backend language (Go, TypeScript, etc.).

  • Strong systems thinking around reliability, latency, cost, and complexity tradeoffs.

  • Experience supporting ML/AI workflows, training pipelines, or evaluation systems.

  • Familiarity with data quality, lineage, observability, and governance tooling.

  • Prior ownership of internal platforms used by many teams.

If you’re excited about this role, we encourage you to apply even if your experience doesn’t match every qualification listed above.

Perplexity

About the company

Perplexity

Startup

Perplexity is an innovative technology company that specializes in developing advanced artificial intelligence solutions aimed at enhancing human-computer interaction. With a focus on natural language processing and machine learning, Perplexity empowers users to access information and insights more intuitively and efficiently. The company is dedicated to creating tools that simplify complex data and foster informed decision-making, thereby transforming the way individuals and organizations engage with knowledge. Through its commitment to excellence and user-centric design, Perplexity is shaping the future of information retrieval and analysis.