Software Engineer - Training/Inference (C++)
xAI · Palo Alto, California
Architect and implement scalable distributed infrastructure for model serving (load balancing, auto-scaling, batch scheduling, global KV cache).; Optimize latency and throughput of model inference under real production…
AI/LLM
This posting is pulled directly from xAI's own Greenhouse page and applying happens on their site, not here. Posted 2026-08-27; Skip The Boards re-checks every company's board regularly and removes listings that disappear from the source.