Role at a glance
- Salary
- $200K – $350K/yr
- Location
- San Francisco, United States New York City, United States
- Work arrangement
- Remote
- Employment
- Full-time
- Experience
- 3+ years of software engineering experience shipping production systems
Spotted an issue?
We’ll check it against the original posting.
Role Summary
The Answer Quality team builds evaluations for Perplexity’s LLM-first search engine, specialized data sources, and models. This role develops the data flywheel and evaluation infrastructure that helps Search, Product, and other teams measure and improve Answer Quality.
What You'll Do
- Build systems and pipelines that provide reliable evaluation verdicts to Search, Product, and other teams
- Turn raw signals into durable datasets for company-wide decision-making
- Build a simulator pipeline that replays user interactions in formats legible to LLMs and VLMs
- Implement monitoring, lineage, and quality checks for downstream data reliability
- Shape how Perplexity measures and improves Answer Quality
Generated from the employer's posting. Verify important details before applying.
View full postingQualifications
Strong proficiency in Python and SQL; experience with big data systems, distributed compute, large-scale storage, data modeling, system design, debugging distributed systems, AWS, and Databricks or Spark.
Required
- 3+ years of software engineering experience shipping production systems
- Python and SQL
- Big data systems, distributed compute, and large-scale storage
- Data modeling, system design, and debugging distributed systems
- AWS
- Databricks or Spark
- Agentic coding workflows and AI-assisted development tools
Preferred
- Data engineering background including pipelines, orchestration, and warehousing patterns
- LLM/VLM interfaces, tokenization, structured formats, and multimodal payloads
- Evaluation platforms, experimentation systems, or machine learning infrastructure
- Supporting customer-facing products at scale
Original job description
Content provided by the employer
Original job description
Content provided by the employer
Perplexity serves tens of millions of users daily with reliable, high-quality answers grounded in an LLM-first search engine and specialized data sources. The Answer Quality team ensures that our prompts, tools, search, and specialized datasets, combined with both frontier and in-house models, create the best possible experience for our users. As our product evolves, our evaluations must remain fast, accurate, and actionable. In this role, you will build the data flywheel that serves teams across Perplexity.
Responsibilities
Build the systems and pipelines that enable Search, Product, and other teams to independently access and utilize reliable eval verdicts without bottlenecks
Take ownership of the "evals-to-product" loop, autonomously determining the best way to turn raw signals into durable datasets that power decision-making across the company
Build a robust simulator pipeline capable of replaying user interactions with the product in formats legible to LLMs and VLMs, reflecting product changes as they are shipped
Maintain data trust by implementing monitoring, lineage, and quality checks, ensuring downstream consumers can rely on the results implicitly
Operate in a small, high-impact team where your work directly shapes how Perplexity measures and improves Answer Quality
Qualifications
3+ years of software engineering experience shipping production systems
Strong proficiency in Python and SQL with the ability to write production-grade, maintainable code
Experience with big data systems including distributed compute and large-scale storage
Solid fundamentals in data modeling, system design, and debugging distributed systems
Experience with AWS and lakehouse ecosystems like Databricks or Spark
Comfortable with agentic coding workflows and using AI-assisted development tools to iterate faster
Preferred Qualifications
Data engineering background including pipelines, orchestration, and warehousing patterns
Familiarity with LLM/VLM interfaces, tokenization, structured formats, and multimodal payloads
Experience with evaluation platforms, experimentation systems, or machine learning infrastructure
Prior work supporting customer-facing products at scale
About the company
Perplexity
Startup
Perplexity is an innovative technology company that specializes in developing advanced artificial intelligence solutions aimed at enhancing human-computer interaction. With a focus on natural language processing and machine learning, Perplexity empowers users to access information and insights more intuitively and efficiently. The company is dedicated to creating tools that simplify complex data and foster informed decision-making, thereby transforming the way individuals and organizations engage with knowledge. Through its commitment to excellence and user-centric design, Perplexity is shaping the future of information retrieval and analysis.
Perplexity is an innovative technology company that specializes in developing advanced artificial intelligence solutions aimed at enhancing human-computer interaction. With a focus on natural language processing and machine learning, Perplexity empowers users to access information and insights more intuitively and efficiently. The company is dedicated to creating tools that simplify complex data and foster informed decision-making, thereby transforming the way individuals and organizations engage with knowledge. Through its commitment to excellence and user-centric design, Perplexity is shaping the future of information retrieval and analysis.