Jobs · empllo
Staff / Principal Machine Learning Engineer, Serving - Switzerland
Lumen Dynamics · Europe · Posted 3d ago
About the role
Lead real-time inference optimization for serving and implement model acceleration techniques.
📋 Description Lead real-time inference optimization for serving Implement model acceleration: quantization, distillation Optimize performance on GPUs (C++, CUDA, Rust, Python) Build scalable distributed systems (Kubernetes, Ray) Own end-to-end deployment from research to prod Collaborate with cross-geo teams and leadership 🎯 Requirements Inference optimization: vLLM or TRT-LLM Model acceleration: quantization, distillation, caching, batching, paged attention High-performance systems: C++, CUDA, Rust, Python; GPU profiling Distributed systems: Kubernetes, Ray, load balancing, multi-GPU/multi-node Public work: OSS contributions to major inference engines Background: PhD in CS/Physics/Math or equivalent experience 🎁 Benefits Flat structure, fast iterations Open-source contributions encouraged Impact-focused culture with visibility of work Remote work within Switzerland
Read the full posting on empllo →
FAQ
Is the Staff / Principal Machine Learning Engineer, Serving - Switzerland role at Lumen Dynamics remote?+
This Staff / Principal Machine Learning Engineer, Serving - Switzerland position is listed as remote (Europe).
What seniority level is this Staff / Principal Machine Learning Engineer, Serving - Switzerland role?+
This is a lead level position.
How do I apply for the Staff / Principal Machine Learning Engineer, Serving - Switzerland role at Lumen Dynamics?+
Use the "Apply on empllo" button to open the original posting on empllo, where you can submit your application directly to Lumen Dynamics.