Member of Technical Staff (Software Engineer, Data Flywheel)
Track this application
Get Started FreeMatch score against your CV
Get Started FreeTailor your resume to this job
Get Started FreeAbout interviewing at Perplexity
One of the few AI startups with a fully published interview guide (perplexity.ai/hub/careers/interview-guide): online application (response within two weeks) → recruiter phone screen → a technical screen that for engineers is 'usually a standard technical programming interview' → a quickly-scheduled onsite of 4–5 interviews including a hiring-manager deep dive on past work and experience anecdotes → a final interview with a Perplexity founder or leader → decision within a week of the onsite. Coding leans Python and mixes LeetCode medium–hard with practical search-flavored tasks (ranking/filtering, concurrency, data handling); system design is AI-native (RAG pipelines, retrieval at scale, LLM serving cost/latency). Applicants are judged 'solely on merit and potential impact' and must show 'frontier knowledge and excellence in at least one area'; roles are broad by default with team matching happening during the onsite, every role — managers included — is hands-on, and building AI products isn't expected but fluency in using AI tools is required. In-person 4 days/week near an office; remote is case-by-case.
Read the full Perplexity interview process →Description
Perplexity serves tens of millions of users daily with reliable, high-quality answers grounded in an LLM-first search engine and specialized data sources. The Answer Quality team ensures that our prompts, tools, search, and specialized datasets, combined with both frontier and in-house models, create the best possible experience for our users. As our product evolves, our evaluations must remain fast, accurate, and actionable. In this role, you will build the data flywheel that serves teams across Perplexity.
Responsibilities
Build the systems and pipelines that enable Search, Product, and other teams to independently access and utilize reliable eval verdicts without bottlenecks
Take ownership of the "evals-to-product" loop, autonomously determining the best way to turn raw signals into durable datasets that power decision-making across the company
Build a robust simulator pipeline capable of replaying user interactions with the product in formats legible to LLMs and VLMs, reflecting product changes as they are shipped
Maintain data trust by implementing monitoring, lineage, and quality checks, ensuring downstream consumers can rely on the results implicitly
Operate in a small, high-impact team where your work directly shapes how Perplexity measures and improves Answer Quality
Qualifications
3+ years of software engineering experience shipping production systems
Strong proficiency in Python and SQL with the ability to write production-grade, maintainable code
Experience with big data systems including distributed compute and large-scale storage
Solid fundamentals in data modeling, system design, and debugging distributed systems
Experience with AWS and lakehouse ecosystems like Databricks or Spark
Comfortable with agentic coding workflows and using AI-assisted development tools to iterate faster
Preferred Qualifications
Data engineering background including pipelines, orchestration, and warehousing patterns
Familiarity with LLM/VLM interfaces, tokenization, structured formats, and multimodal payloads
Experience with evaluation platforms, experimentation systems, or machine learning infrastructure
Prior work supporting customer-facing products at scale
Salary Range: $200K - $350K