CA$250,000 – CA$535,000
Listed on Cohere’s own careers site. You apply with them directly — we never stand between you and the employer.
What this role is
Cohere seeks an engineer to optimize audio model inference performance across latency, throughput, and quality metrics. This role suits someone with deep systems expertise in machine learning inference who can identify bottlenecks and deliver solutions for real-time audio processing at scale.
Our summary, not Cohere’s wording. The full posting is on their site.
Skills this role names
Log in to see which of these are already on your profile.
What they ask for
Required
- High-performance audio or machine learning inference systems development
- C++ and Python proficiency
- Deep learning models for audio, speech, or language applications
- Results-oriented mindset
Nice to have
- GPU programming and low-level system optimization
- Model parallelization across multiple GPUs
- Duplex real-time streaming architectures
- Machine learning framework internals for audio
- Custom distributed inference systems
- Transformers and sequence modeling for audio/speech
- End-to-end audio pipeline optimization