The opportunity
Fin is the AI Customer Agent company on a mission to help businesses provide perfect customer experiences.
What you'll do
A track record of working on model training or model inference at scale , or: on low‑level GPU coding (e.g. CUDA, Triton). Experience with one is great, multiple is even better.
Implement and scale training pipelines for large transformer and LLM models,: from data ingestion and preprocessing through distributed training and evaluation.
Build and optimize inference services that deliver low‑latency,: high‑reliability experiences for our customers, including autoscaling, routing, and fallbacks.
Work on GPU‑level performance: tuning kernels, improving utilization, and identifying bottlenecks across our training and inference stack.
Collaborate closely with ML scientists to implement cutting edge training and: inference methods and bring them to production.
Play an active role in hiring, mentoring, and developing other engineers on the team.
What they're looking for
- Raise the bar for technical standards, reliability, and operational excellence across Fin's AI platform.
- You have 5+ years of experience in software engineering , with a strong track: record of shipping high‑quality products or platforms.
- You hold a degree in Computer Science, Computer Engineering, or a related: field (or you have equivalent experience with very strong fundamentals).
- You have hands‑on experience with one or more of the following: