I

Staff / Principal Machine Learning Engineer, Serving - Switzerland

Inworld

, , switzerland, , , switzerland, Switzerland Full-time June 13, 2026
Apply Now

Vacancy Description

Who We're Looking For

A year ago, reliably working agentic systems and sub‑second multimodal inference at scale barely existed. Nobody has a decade of experience here. So we're not screening for a resume template — we're looking for strong people from varied backgrounds who learn fast, thrive in ambiguity, and can show us what they've built, broken, and understood.

Experience We Find Useful

  • Inference Optimization. Deep understanding of modern serving frameworks and techniques like vLLM or TRT‑LLM.

  • Model Acceleration . Hands‑on experience with quantization, distillation, caching strategies, continuous batching, paged attention, and speculative decoding.

  • High‑Performance Systems. Proficiency in C++, CUDA, Rust, or highly optimised Python. You know how to profile code and squeeze every ounce of performance out of NVIDIA GPUs.

  • Distributed Systems & Scaling. Experi...

Ready to Apply?

अभी आवेदन करें

Submit your application for Staff / Principal Machine Learning Engineer, Serving - Switzerland at Inworld

Apply for this Position