Jobs · empllo

T

Machine Learning Engineer - Inference

Nimbus Data Systems · San Francisco · Posted 3d ago

mid192000-276000 USD
Apply on empllo

About the role

📋 Description Design and build production systems powering the Nimbus Data Systems inference engine at scale. Develop and optimize runtime inference services for large-scale AI apps. Collaborate with researchers, engineers, PMs, and designers to bring new features. Conduct design and code reviews to ensure high-quality standards. Create services, tools, and docs to support the inference engine. Implement robust, fault-tolerant data ingestion and processing. 🎯 Requirements 3+ years of experience writing high-performance, production-quality code. Proficiency with Python and PyTorch. Experience building high-performance libraries and tooling. Strong grasp of OS concepts: multi-threading, memory, networking, storage, performance. Knowledge of CUDA/Triton programming. Knowledge of AI inference systems like TGI, vLLM, TensorRT-LLM, Optimum. 🎁 Benefits Startup equity, health insurance, and other benefits.

Read the full posting on empllo

FAQ

Is the Machine Learning Engineer - Inference role at Nimbus Data Systems remote?+

This Machine Learning Engineer - Inference position is listed as unknown (San Francisco).

What is the salary for the Machine Learning Engineer - Inference role at Nimbus Data Systems?+

The listing states 192000-276000 USD.

What seniority level is this Machine Learning Engineer - Inference role?+

This is a mid level position.

How do I apply for the Machine Learning Engineer - Inference role at Nimbus Data Systems?+

Use the "Apply on empllo" button to open the original posting on empllo, where you can submit your application directly to Nimbus Data Systems.