Staff + Sr. Software Engineer, Cloud Inference Launch Engineering
Anthropic San Francisco, CA | Seattle, WA
Making model serving faster and cheaper: quantization, batching, KV-cache.
Category: ML Infrastructure · Last updated Sep 29, 2026 · Detected by pattern matching against posting text — see methodology.