Member of Technical Staff (Software Engineer, Data Flywheel)
Full Description
Perplexity serves tens of millions of users daily with reliable, high-quality answers grounded in an LLM-first search engine and specialized data sources. The Answer Quality team ensures that our prompts, tools, search, and specialized datasets, combined with both frontier and in-house models, create the best possible experience for our users. As our product evolves, our evaluations must remain fast, accurate, and actionable. In this role, you will build the data flywheel that serves teams across Perplexity. RESPONSIBILITIES - Build the systems and pipelines that enable Search, Product, and other teams to independently access and utilize reliable eval verdicts without bottlenecks - Take ownership of the "evals-to-product" loop, autonomously determining the best way to turn raw signals into durable datasets that power decision-making across the company - Build a robust simulator pipeline capable of replaying user interactions with the product in formats legible to LLMs and VLMs, reflecting product changes as they are shipped - Maintain data trust by implementing monitoring, lineage, and quality checks, ensuring downstream consumers can rely on the results implicitly - Operate in a small, high-impact team where your work directly shapes how Perplexity measures and improves Answer Quality QUALIFICATIONS - 3+ years of software engineering experience shipping production systems - Strong proficiency in Python and SQL with the ability to write production-grade, maintainable code - Experience with big data systems including distributed compute and large-scale storage - Solid fundamentals in data modeling, system design, and debugging distributed systems - Experience with AWS and lakehouse ecosystems like Databricks or Spark - Comfortable with agentic coding workflows and using AI-assisted development tools to iterate faster PREFERRED QUALIFICATIONS - Data engineering background including pipelines, orchestration, and warehousing patterns - Familiarity with LLM/VLM interfaces, tokenization, structured formats, and multimodal payloads - Experience with evaluation platforms, experimentation systems, or machine learning infrastructure - Prior work supporting customer-facing products at scale
Why This Role?
Bekerja di tim kecil dengan dampak tinggi, di mana pekerjaan Anda langsung mempengaruhi kualitas jawaban Perplexity.
Key Responsibilities
- Bangun sistem dan pipa data yang memungkinkan tim untuk mengakses hasil evaluasi yang tepercaya tanpa hambatan
- Ambil alih loop 'evals-to-product', menentukan cara terbaik untuk mengubah sinyal mentah menjadi dataset yang kuat
- Bangun pipa simulator yang mampu mereplay interaksi pengguna dengan produk dalam format yang dapat dibaca oleh LLM dan VLM
- Pertahankan kepercayaan data dengan mengimplementasikan monitoring, garis keturunan, dan pemeriksaan kualitas
- Bekerja di tim kecil dengan dampak tinggi, di mana pekerjaan Anda langsung mempengaruhi kualitas jawaban Perplexity
Requirements
- Punya pengalaman 3+ tahun sebagai software engineer dalam mengirimkan sistem produksi
- Mampu menulis kode Python dan SQL yang berkualitas produksi dan dapat dipertahankan
- Punya pengalaman dengan sistem data besar termasuk komputasi terdistribusi dan penyimpanan skala besar
- Punya dasar yang kuat dalam pemodelan data, desain sistem, dan debugging sistem terdistribusi
- Punya pengalaman dengan AWS dan ekosistem lakehouse seperti Databricks atau Spark
Required Skills
Keywords
View Original Description from Ashby Job Boards
Original description from Ashby Job Boards
Perplexity serves tens of millions of users daily with reliable, high-quality answers grounded in an LLM-first search engine and specialized data sources. The Answer Quality team ensures that our prompts, tools, search, and specialized datasets, combined with both frontier and in-house models, create the best possible experience for our users. As our product evolves, our evaluations must remain fast, accurate, and actionable. In this role, you will build the data flywheel that serves teams across Perplexity. RESPONSIBILITIES - Build the systems and pipelines that enable Search, Product, and other teams to independently access and utilize reliable eval verdicts without bottlenecks - Take ownership of the "evals-to-product" loop, autonomously determining the best way to turn raw signals into durable datasets that power decision-making across the company - Build a robust simulator pipeline capable of replaying user interactions with the product in formats legible to LLMs and VLMs, reflecting product changes as they are shipped - Maintain data trust by implementing monitoring, lineage, and quality checks, ensuring downstream consumers can rely on the results implicitly - Operate in a small, high-impact team where your work directly shapes how Perplexity measures and improves Answer Quality QUALIFICATIONS - 3+ years of software engineering experience shipping production systems - Strong proficiency in Python and SQL with the ability to write production-grade, maintainable code - Experience with big data systems including distributed compute and large-scale storage - Solid fundamentals in data modeling, system design, and debugging distributed systems - Experience with AWS and lakehouse ecosystems like Databricks or Spark - Comfortable with agentic coding workflows and using AI-assisted development tools to iterate faster PREFERRED QUALIFICATIONS - Data engineering background including pipelines, orchestration, and warehousing patterns - Familiarity with LLM/VLM interfaces, tokenization, structured formats, and multimodal payloads - Experience with evaluation platforms, experimentation systems, or machine learning infrastructure - Prior work supporting customer-facing products at scale
Free account · no credit card · Log in
Pro $9/mo · unlimited applies + AI resume
Explore related
Market data & reports
Salary & skill-demand research built from our own listings data.
- Indonesia IT Jobs vs Global Remote (2026)Primary analysis of 2,049 listings: methodology, classification rules, downloadable datasets.
- AI-Skill Demand: Indonesia vs Global Remote (2026)10,000+ postings, taxonomy-first classifier, Wilson CIs, pre-registered before analysis.
- Indonesia Hiring Report: Tech vs Non-TechJob demand by field from aggregate open-job counts — never individual listings.
- Indonesia Salary BenchmarkAggregate salary ranges across roles, with open methodology and dataset.
- Indonesian Remote Work Salary & Demand IndexHow much of the global remote job corpus is open to Indonesia, and what it pays (USD) by role.
- Indonesia Quarterly Labor Market ReportLayoffs, funding, salaries & skills per quarter — open aggregates.
- Remote Market Reports by RoleAuto-generated per role family — skills, seniority, companies, salary.
- Global Remote Salary BenchmarkAnnual salary by role & currency, plus the share of listings open worldwide.
Free account · no credit card · Log in
Pro $9/mo · unlimited applies + AI resume