A pioneering technology company, at the forefront of innovation in machine learning and artificial intelligence, is seeking a Principal Machine Learning Engineer. This organization is renowned for its deep technical expertise and commitment to pushing the boundaries of what AI can achieve, working on groundbreaking projects that redefine industries.
The Role
- Architect and build large-scale ML systems across the entire lifecycle, from data to deployment.
- Design and optimize high-performance training pipelines utilizing GPU infrastructure.
- Architect inference systems that expertly balance latency, throughput, cost, and reliability at scale.
- Implement and maintain robust data systems for both synthetic and real-world training data.
- Own production deployment strategies, including GPU optimization and scaling policies.
- Collaborate closely with application engineering to seamlessly integrate ML systems into diverse products.
What You'll Need
- Strong background in deep learning and transformer-based architectures with AI experience.
- Hands-on experience training, fine-tuning, or deploying large-scale ML models in production environments.
- Proficiency with modern ML frameworks (e.g., PyTorch, JAX) and adaptability to new technologies.
- Experience with distributed training and inference frameworks (e.g., DeepSpeed, FSDP, Megatron, ZeRO, Ray).
- Solid software engineering fundamentals, capable of building robust and maintainable production-grade systems.
- Demonstrated experience with GPU optimization, memory efficiency, and large-scale data processing (e.g., Apache Arrow, Spark).
What's On Offer
- The opportunity to design and evolve critical ML systems in a hands-on, high-impact role.
- A fully remote position allowing for significant flexibility.
- Comprehensive benefits package including medical, dental, vision, savings plans, and PTO.
Apply via Haystack today!