The role
- COMP
- $230K - $350K
- EQUITY
- Competitive
- LOCATION
- South Bay Area
- WORKPLACE
- On-site
- EXPERIENCE
- 2 - 10 years
- VISA
- None, Visa transfers, New visa sponsorships
- STACK
- Rust, Python, PyTorch, C/C++, Go
- INDUSTRY
- AI
/roles — ROLE_319
Stealth inference startup building a new LLM serving engine in Rust from first principles
Stealth-stage company building a high-performance LLM inference platform in Rust, focused on raw speed and serving efficiency, led by founders with deep AI infrastructure experience.
JD — the work
You would help build an LLM inference platform from scratch, written primarily in Rust and with no legacy code to work around. The role suits someone with a systems background who knows inference internals such as KV caching, attention, batching, and scheduling, and who wants to own the whole serving stack rather than a narrow slice. The company is small, backed by well-known investors, founded by engineers with deep AI infrastructure backgrounds, and still in stealth ahead of announcing its funding and product.