This role is fully remote but the employer can only hire in San Francisco. Make sure you're eligible to work there before applying. See work-from-anywhere jobs →
About this role
We build and run the inference engine behind every Perplexity query and deploy dozens of model architectures at scale with tight latency and cost budgets. Our stack is Rust, Python, CUDA, and CuTe DSL
Found this role here? Please mention AnywhereJobs in your application — it helps us keep surfacing truly location-independent jobs.
Scam safety
You should never have to pay to apply for a job, and a legitimate employer will never ask you to buy your own equipment upfront or share bank details before hiring. If a listing asks for money, report it and walk away.