The opportunity
Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems.
What you'll do
Stay up-to-date with the latest research in code LLMs, agents, and related: fields, implementing novel ideas into our systems.
Design and implement scalable strategies to train code models, and deploy: agent frameworks for inference and sampling. You will be collaborating with the pretraining team, create SFT trajectories and work on existing and new RL algorithms
Hillclimb on existing benchmarks and design new ones that reflect the needs of our enterprise users
Lead experiments on our state-of-the-art compute infrastructure, pushing the: boundaries of what’s possible with frontier LLMs.
A PhD in Computer Science, Machine Learning, or a related field, with: publications in top-tier venues (e.g., NeurIPS, ICML, ICLR, ACL, EMNLP).
Deep expertise in code LLMs and agent systems, with a strong understanding of: the latest research and trends. We are looking for people who not only have worked with code models, but have actively contributed to their development
What they're looking for
- Hands-on experience with frontier LLMs and their applications in code generation or automation.
- Strong software engineering skills, with proficiency in Python and PyTorch, TensorFlow, or similar frameworks.
- Experience with distributed systems, cloud infrastructure, and scalable architectures.
- A proactive, self-motivated mindset, with a passion for solving ambitious, open-ended problems.