Pubblica annuncio gratuito
Ricevi nuove offerte di lavoro

Crea una Job Alert gratuita per machine learning engineer, vla & rl - senior / milano

Machine Learning Engineer, VLA & RL - Senior

Pubblicato il 09-08-2026 - Cyberwave in Milano

Overview

Machine Learning Engineer, VLA & RL - Senior. Full-time. TC starting at 70k. Posted 1 day ago.

Build vision-language-action and reinforcement learning models for real-world robotic systems. Train policies that generalize across embodiments, tasks, simulators, and physical deployments.

About Cyberwave

Cyberwave is building the infrastructure layer for intelligent machines - making robotics as accessible, scalable, and programmable as cloud software. Our platform connects simulation, digital twins, edge devices, cloud training, and real robots into one operating layer for robotics teams.

Role

We're looking for a Machine Learning Engineer focused on vision-language-action (VLA) models, reinforcement learning, and cross-embodiment transfer. You'll work on models that turn perception, language, and task context into robot actions across different hardware platforms: arms, mobile robots, drones, and other industrial systems.

This is a hands-on applied ML role. We care about candidates who have trained and evaluated real policies, debugged failures across simulation and hardware, and understand the gap between promising demos and reliable deployment. You'll work closely with robotics, simulation, infrastructure, and product teams to build learning systems that can be trained at scale, evaluated rigorously, and deployed safely on real robots.

This role is based in Milan or Zurich, with regular access to real robots, simulation infrastructure, and customer-facing deployment scenarios.

Work Style

Hands-on applied ML for embodied AI, simulation, and real robot deployments

Requirements





- 3+ years of hands-on experience building ML systems for robotics, embodied AI, reinforcement learning, or visuomotor control
- Specific experience with vision-language-action (VLA) models, robotic foundation models, imitation learning, behavior cloning, or language-conditioned policies
- Strong experience with reinforcement learning algorithms and workflows, such as PPO, SAC, offline RL, RL fine-tuning, reward modeling, or policy evaluation
- Practical experience with cross-embodiment transfer, including transferring policies across robot morphologies, sensors, action spaces, simulators, or real hardware platforms
- Experience training and evaluating policies in simulation environments such as MuJoCo, Isaac Sim/Lab, PyBullet, ManiSkill, robosuite, or similar robotics simulators
- Strong Python and PyTorch skills, with good software engineering habits for reproducible training, experiment tracking, datasets, and evaluation
- Comfort debugging model failures across perception, action representations, control loops, latency, data quality, and hardware behavior
- Comfortable working in English in an international, fast-moving environment

Responsibilities





- Train and evaluate VLA, imitation learning, and reinforcement learning policies for real robotic tasks
- Build model and data pipelines for language-conditioned robot control, visuomotor policies, trajectory datasets, and action representations
- Design experiments for cross-embodiment transfer across arms, mobile robots, drones, simulated systems, and physical hardware
- Improve policy robustness through simulation, domain randomization, dataset curation, offline evaluation, online rollouts, and sim-to-real validation
- Collaborate with robotics and infrastructure teams to deploy learned policies into Cyberwave's edge, simulation, and digital twin stack
- Create rigorous evaluation suites for task success, generalization, safety, latency, and real-world reliability
- Stay close to frontier research in embodied AI while turning useful ideas into production-quality systems

What We Offer

Work on frontier embodied AI with direct paths to real robot deployment

Access to real robots, simulation infrastructure, and robotics datasets

Competitive compensation and meaningful equity

Join a high-talent team of repeat founders, ex-Google engineers, and PhDs

In-person collaboration in Milan and Zurich with real hardware and high ownership

Ready to Join Our Team?

We'd love to hear from you! When applying, please include:

- Your Github or LinkedIn profile
- 2-3 lines about why you would like to join Cyberwave

Tell us what excites you about this opportunity and how you can contribute to our mission!

#J-18808-Ljbffr

» RISPONDI A QUESTO ANNUNCIO
Altri Annunci
Consulente del credito per le PMI - ROMAGNA
2026-08-10 00:19:10 - Pmitutoring.It - Roma
CONSULENTE DEL CREDITO PER LE IMPRESE zona ROMAGNA Per ampliare la propria rete vendita, PMI Tutoring by BFS Partner S.p.A., mediatore creditizio per le PMI, cerca nuovi consulenti del credito per larea della Romagn [...]
Consulente Del Credito - Cesena
2026-08-10 00:19:03 - Euroansa - Cesena
Consulente del Credito - Euroansa Costruisci il tuo futuro nel mondo della consulenza creditizia Vuoi entrare in un settore solido, meritocratico e in continua crescita? Con Euroansa puoi trasformare ambizione e t [...]
Ricevi nuove offerte di lavoro

Crea una Job Alert gratuita per machine learning engineer, vla & rl - senior / milano

Consulente Del Lavoro
2026-08-10 00:19:02 - Iqm Selezione S. R. L - Parma
IQM Selezione, società di head-hunting che da 20 anni accompagna Aziende e Professionisti nella crescita con un approccio etico, consulenziale e fortemente attento al valore delle persone, per importante realtà pr [...]
Consulente Del Credito P.Iva (Roma)
2026-08-10 00:18:59 - Gruppo Mol - Roma
In Mavriq , parte di Moltiply Group, aiutiamo le persone a fare la scelta giusta - che si tratti di un mutuo, un’assicurazione, una fornitura di energia o un piano telefonico. I nostri brand, tra cui MutuiOnlin [...]