The opportunity
Meet DeepL DeepL is a global AI product and research company focused on building secure, intelligent solutions to complex business problems. Over 200,000 business customers and millions of individuals across 228 global markets today trust DeepL's Language AI platform for…
What they're looking for
- Mentor researchers and engineers, promote hands-on collaboration, and raise the bar for model quality.
- Proven experience making large models steerable and instruction-following by: identifying the most effective method to instill a given behavior, drawing from instruction tuning, latent space methods, steering vectors, and/or constrained encoding and decoding methods.
- Deep, hands-on expertise in LLM post-training (SFT, DPO), knowledge: distillation (teacher-student training), and/or reinforcement learning (RLHF/RLAIF, PPO/GSPO, and reward modeling).
- Strong data-centric instincts for building synthetic-data and preference-data: pipelines, LLM-as-judge generation, data curation and filtering, and reasoning about data mixtures and ablations.