|
Salary
usd 200,000 - 2..
|
Remote
Location
πΊπΈ
USA Only
|
|
Employment Type
full-time
|
Posted
4d ago
|
4d ago - Cantina is hiring a remote Machine Learning Engineer, Speech - Joint Audio-Video Modeling. πΈ Salary: usd 200,000 - 220,000 per year πLocation: USA
Role Description
We're looking for a Research / ML Engineer to join our Speech Team to build state-of-the-art speech and audio generation systems end-to-end from data specs through production inference with a focus on joint audio-video modeling.
You'll own the audio side of multimodal generation:
This includes:
You'll drive the model β data β eval flywheel, partnering closely with research, video, data, and infra to ship fast, reliable, and cost-aware models. In this role you'll work at the intersection of cutting-edge research and practical engineering, contributing to the development of safe, steerable, and trustworthy AI systems.
You will thrive in this role if you:
Qualifications
Requirements
Benefits
| πΊπΈ | Be aware of the location restriction for this remote position: USA Only |
| βΌ | Beware of scams! When applying for jobs, you should NEVER have to pay anything. Learn more. | οΈ
|
Salary
usd 200,000 - 2..
|
Remote
Location
πΊπΈ
USA Only
|
|
Employment Type
full-time
|
Posted
4d ago
|
| πΊπΈ | Be aware of the location restriction for this remote position: USA Only |
| βΌ | Beware of scams! When applying for jobs, you should NEVER have to pay anything. Learn more. | οΈ
Access 125,000+ vetted remote jobs and get daily alerts.