Member of Technical Staff (Data Scientist, Evals)

Perplexity
San Francisco, Palo Alto
Workplace: HybridFull timeUSD 200,000 - 300,000 annuallyFunction: Data Science & Machine LearningExperience: 4+ yearsEducation: mastersSkills: ["Python","SQL","Machine learning","Data science","AI"]

We are seeking a data scientist to build and maintain automated evaluation pipelines for Perplexity’s LLM-first search platform. You will design evaluation sets, develop VLM-based renderings, and measure answer quality across products, collaborating with leadership to translate metrics into product improvements.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Perplexity
Perplexity
6 months ago

Member of Technical Staff (Data Scientist, Evals)

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 9 hours agoStatus: Live

Job Summary

We are seeking a data scientist to build and maintain automated evaluation pipelines for Perplexity’s LLM-first search platform. You will design evaluation sets, develop VLM-based renderings, and measure answer quality across products, collaborating with leadership to translate metrics into product improvements.
Location: San Francisco, Palo Alto
Workplace: Hybrid
Employment Type: Full time
Job Function: Data Science & Machine Learning

Key Responsibilities

  • •Architect and maintain automated evaluation pipelines to assess answer quality across Perplexity's products, ensuring high standards for accuracy and helpfulness.
  • •Design evaluation sets and methods specifically to measure the impact of tool calls (particularly web search retrieval) on the final answer's quality.
  • •Develop VLM-based solutions to programmatically evaluate how final answers render visually across different platforms and devices.
  • •Continuously review public benchmarks and academic evaluations for their applicability to the Perplexity product, adapting and incorporating them into our regular performance measurements.
  • •Operate within a small, high-impact team where your evaluation metrics directly shape product changes, collaborating closely with technical leadership to measure and improve Answer Quality

Pay and Benefits

Salary: USD 200,000 - 300,000 annually

Key Requirements

  • •PhD or MS in a technical field or equivalent experience.
  • •4+ years of experience in data science or machine learning.
  • •Strong proficiency in Python and SQL (production-grade code).
  • •Experience building within a modern cloud data stack, specifically AWS and Databricks.
  • •Comfortable with agentic coding workflows and using AI-assisted development tools to iterate faster.
Experience:4+ yearsLLMsEvaluationData scienceMachine learningAI
Education:Master's
Skills:PythonSQLMachine learningData scienceAI
Languages:English
Tech Stack:PythonSQLAWSDatabricksLLMsVLMs

Company Brief

Perplexity
Perplexity AI provides an AI-powered answer engine that returns conversational, citation-backed answers to user queries and offers Pro/Enterprise products and APIs for research, knowledge work, and search augmentation.
Industry: AI & Machine Learning
Company Size: Medium (51 to 250 employees)
Revenue: USD 50M to 100M
Growth: Scaleup
Valuation: Unicorn (USD 1B+)
Funding: Series D
Headquarters: San Francisco, United States
Founded: 2022
Glassdoor
Glassdoor: 4.6
WebsiteLinkedInGlassdoor