Member of Technical Staff (TPM, Inference) at perplexity
Company: perplexity
Location: San Francisco
Workplace: On-site
Seniority: Staff+
Category: Data · AI Research
Compensation: Not listed: salary undisclosed on source
Posted: 2026-09-02
Stack: llm
Role overview
WHAT YOU'LL DO - Execute the roadmap for the inference platform — request handling, rate limits and quotas, usage controls, and the reliability and observability surface engineering and product teams depend on - Be the connective tissue between model providers and Perplexity's engineering and product teams — coordinating onboarding, launch readiness, and rollout for new models and capacity - Drive latency, throughput, uptime, and cost-efficiency as core execution metrics,…