Back to jobs

Member of Technical Staff (Search Quality Analyst)

Perplexity
Belgrade (Hybrid)
Full-time
You apply on Perplexity's own careers site

This role improves Perplexity’s search and answer systems by diagnosing quality issues, defining metrics, and building evaluation and training datasets. You’ll work across data analysis and engineering, with your findings informing search algorithms and product experiments.

Permanent
Hybrid
4+ years

Skills & Expertise

Python
SQL
Search quality analysis
Metric design
LLM-as-a-judge labeling pipelines
A/B testing
Apache Spark
Databricks

Key Responsibilities

Diagnose quality issues in the search pipeline and improve search snippets and page selection.

Design search quality metrics, datasets, and labeling pipelines for training and evaluation.

Run A/B experiments to validate search and answer system improvements.

Full Description

Perplexity is looking for an experienced analyst to help us build and improve our core search technologies. You'll work at the intersection of data analysis and engineering - designing metrics, building data pipelines, and improving the quality of our search and answer systems.

This role is hybrid in Belgrade, London or Berlin.

Responsibilities

• Find and diagnose quality issues in our search pipeline

• Design metrics from scratch to track and measure search quality

• Build datasets for model training, including LLM-as-a-judge labeling pipelines

• Improve search snippet quality and page selection algorithms for indexing

• Design and analyze A/B experiments to validate improvements

Qualifications

• 4+ years of experience as a data analyst or in a related role

• Strong coding skills — expected to write production-grade code at a mid-level backend engineer level

• Proficiency with SQL and Python

• Demonstrated hands-on experience with at least one of the following:

• Designing metrics from scratch (not just analyzing existing A/B experiments)

• Building labeling pipelines using LLM-as-a-judge

• Training ML models that shipped to production with measurable metric improvements

• Designing evals with known ground truth (e.g. SimpleQA, BrowseComp) or driving meaningful improvements on such evals

Preferred Qualifications

• Experience working on search-related products

• Knowledge of statistics and A/B experiment design

• Experience with Apache Spark or Databricks

Applications are handled on Perplexity's site