Jobs · empllo
Machine Learning Engineer - Inference
Nimbus Data Systems · San Francisco · Posted 3d ago
About the role
📋 Description Design and build production systems powering the Nimbus Data Systems inference engine at scale. Develop and optimize runtime inference services for large-scale AI apps. Collaborate with researchers, engineers, PMs, and designers to bring new features. Conduct design and code reviews to ensure high-quality standards. Create services, tools, and docs to support the inference engine. Implement robust, fault-tolerant data ingestion and processing. 🎯 Requirements 3+ years of experience writing high-performance, production-quality code. Proficiency with Python and PyTorch. Experience building high-performance libraries and tooling. Strong grasp of OS concepts: multi-threading, memory, networking, storage, performance. Knowledge of CUDA/Triton programming. Knowledge of AI inference systems like TGI, vLLM, TensorRT-LLM, Optimum. 🎁 Benefits Startup equity, health insurance, and other benefits.
Read the full posting on empllo →
FAQ
Is the Machine Learning Engineer - Inference role at Nimbus Data Systems remote?+
This Machine Learning Engineer - Inference position is listed as unknown (San Francisco).
What is the salary for the Machine Learning Engineer - Inference role at Nimbus Data Systems?+
The listing states 192000-276000 USD.
What seniority level is this Machine Learning Engineer - Inference role?+
This is a mid level position.
How do I apply for the Machine Learning Engineer - Inference role at Nimbus Data Systems?+
Use the "Apply on empllo" button to open the original posting on empllo, where you can submit your application directly to Nimbus Data Systems.