Perplexity Verified 44h ago
Member of Technical Staff (AI Inference Engineer)
Palo Alto, California, United StatesAll locationsPalo Alto, California, United StatesSan Francisco, California, United StatesNew York City, New York, United States On-site
$220,000–$485,000 a yearPay$220,000–$485,000
TypeFull-time
Work settingOn-site
Verified listing
JobFig found this opening at its original source and checks that it remains available.
About the role
You own problems end-to-end. You can read a research paper on Monday, write a kernel on Wednesday, and debug a production incident on Friday.
What you'll bring
- 3+ years of professional software engineering experience with meaningful work on ML inference or high-performance systems.
- Familiarity with at least one deep learning framework (PyTorch, JAX, TensorFlow).
- Understanding of GPU architectures (memory hierarchy, warp scheduling, tensor cores).
- Understanding of common LLM architectures and inference optimization techniques (e.g. quantization, speculative decoding, prefill-decode disaggregation).
- Any other deep systems programming experience is a plus.
Locations
Palo Alto, CaliforniaSan Francisco, CaliforniaNew York City, New York