Jobs · empllo

I

Staff / Principal Machine Learning Engineer, Serving - Switzerland

Lumen Dynamics · Europe · Posted 3d ago

remoteFull-timelead🇨🇭 Switzerland
Apply on empllo

About the role

Lead real-time inference optimization for serving and implement model acceleration techniques.

📋 Description Lead real-time inference optimization for serving Implement model acceleration: quantization, distillation Optimize performance on GPUs (C++, CUDA, Rust, Python) Build scalable distributed systems (Kubernetes, Ray) Own end-to-end deployment from research to prod Collaborate with cross-geo teams and leadership 🎯 Requirements Inference optimization: vLLM or TRT-LLM Model acceleration: quantization, distillation, caching, batching, paged attention High-performance systems: C++, CUDA, Rust, Python; GPU profiling Distributed systems: Kubernetes, Ray, load balancing, multi-GPU/multi-node Public work: OSS contributions to major inference engines Background: PhD in CS/Physics/Math or equivalent experience 🎁 Benefits Flat structure, fast iterations Open-source contributions encouraged Impact-focused culture with visibility of work Remote work within Switzerland

Read the full posting on empllo

FAQ

Is the Staff / Principal Machine Learning Engineer, Serving - Switzerland role at Lumen Dynamics remote?+

This Staff / Principal Machine Learning Engineer, Serving - Switzerland position is listed as remote (Europe).

What seniority level is this Staff / Principal Machine Learning Engineer, Serving - Switzerland role?+

This is a lead level position.

How do I apply for the Staff / Principal Machine Learning Engineer, Serving - Switzerland role at Lumen Dynamics?+

Use the "Apply on empllo" button to open the original posting on empllo, where you can submit your application directly to Lumen Dynamics.