|
Salary
unspecified
|
Remote
Location
|
|
Employment Type
full-time
|
Posted
Today
|
Today - Luma is hiring a remote Research Scientist / Engineer – Reinforcement Learning Infrastructure. 💸 Salary: unspecified 📍Location: Europe
Role Description
You'll build the systems that make reinforcement learning work at frontier scale — coupling policy optimization with large fleets of inference workers, agentic environments, and the reward and verification systems that turn model behavior into learning signal. RL is how Luma's models go from capable to useful.
RL at scale is a full-loop systems problem: training, rollout generation, environment execution, and reward computation running concurrently across thousands of GPUs, all needing to stay fast, stable, and correct together. It fits someone who has lived this — post-trained LLMs with RL, built environments and verifiers, and debugged asynchronous rollout pipelines at scale. If you haven't operated RL at real scale, this will be deep water.
Qualifications
Nice to Have
|
|
Be aware of the location restriction for this remote position: Europe |
| ‼ | Beware of scams! When applying for jobs, you should NEVER have to pay anything. Learn more. | ️
|
Salary
unspecified
|
Remote
Location
|
|
Employment Type
full-time
|
Posted
Today
|
|
|
Be aware of the location restriction for this remote position: Europe |
| ‼ | Beware of scams! When applying for jobs, you should NEVER have to pay anything. Learn more. | ️
Access 130,000+ vetted remote jobs and get daily alerts.
⚡ 130,328+ remote jobs, refreshed hourly
🔔 Real-time alerts: Apply first, direct to employer
🛡️ Vetted companies, no scams, true remote only