Technical Lead, Machine Learning
1 мес. назад
EuropeWorldwideLeadRemote
machine learninggpullmfine-tuningdata pipelinesinferenceevaluationinfrastructure
Lead the development of scalable production machine learning systems for AI-native productivity applications as part of the founding engineering team.
О компании
- We`re partnering with A1 , a new AI venture incubated by BJAK —one of Southeast Asia's largest and profitable technology companies. Backed by an initial $100M investment , A1 is building the next generation of AI-native productivity applications.
- The first product reimagines email using autonomous AI agents that can understand context, reason through multi-step tasks, and complete work on behalf of users while keeping them in control.
- This is an opportunity to join the founding engineering team before product launch and help shape both the platform and the engineering culture.
Обязанности
- We're looking for a Technical Lead, Machine Learning to turn cutting-edge AI research into reliable production systems.
- This isn't a research role. You'll own the infrastructure and execution layer that makes large language models trainable, deployable, observable and scalable in production.
- You'll work across model fine-tuning, inference, evaluation, data pipelines and GPU infrastructure while partnering closely with product and backend engineers.
Требования
- Experience building and shipping production ML systems used by real users.
- Strong Python and PyTorch (or JAX) experience.
- Hands-on experience with LLM fine-tuning and inference optimization.
- Deep understanding of ML infrastructure, distributed training or GPU-based systems.
- Strong engineering mindset with a focus on scalability, reliability and code quality.
- Comfortable taking ownership in a fast-moving startup environment.
Условия
- Founding engineering role with significant technical ownership.
- Opportunity to influence product architecture from the ground up.
- Backed by $100M while operating with the speed of a startup.
- Small, high-talent engineering team focused on solving challenging AI infrastructure problems.
- Competitive cash compensation plus equity.
- Remote-first environment with long-term growth opportunities.
Другое
- Design and scale production ML systems for LLM-based applications.
- Build training and evaluation pipelines for continuous model improvement.
- Fine-tune foundation models using modern adaptation techniques such as LoRA, QLoRA, SFT and DPO.
- Optimize inference performance, latency and GPU utilization.
- Develop high-quality training datasets and evaluation frameworks.
- Own deployment, monitoring and reliability of ML services in production.
- Help define engineering standards while mentoring a small, high-calibre ML team.