Member of Technical Staff (AI Inference Engineer)

Perplexity AI · Palo Alto, California; San Francisco, California; New York, New York · Posted 2026-04-13

Apply on Perplexity AI's site

We build and run the inference engine behind every Perplexity query and deploy dozens of model architectures at scale with tight latency and cost budgets. Our stack is Rust, Python, CUDA, and CuTe DSL - and we need another engineer to join us.

What you will work on ---------------------

Examples of real work the team does:

Who we're looking for ---------------------

Good if you touched any of --------------------------

Qualifications --------------

About Perplexity AI

Perplexity AI is an AI-chat-based conversational search engine that delivers answers to questions using language models.

More jobs at Perplexity AI

Related searches

Updated 2026-10-11.