Job Description
We are pioneering the technological landscape of 2026. As a Senior AI Infrastructure Engineer, you will be at the forefront of building the next generation of scalable, high-performance AI systems. We are looking for a visionary engineer who thrives in a fast-paced environment and is ready to architect the backbone of our machine learning ecosystem.
Why Join Us?
You will work with cutting-edge technology, contribute to open-source frameworks, and have a direct impact on the future of artificial intelligence. We offer competitive compensation, remote flexibility, and a culture that champions innovation.
Responsibilities
- Design and implement scalable machine learning infrastructure and pipelines using Kubernetes, Docker, and cloud-native technologies.
- Optimize model training and inference performance to handle massive datasets and real-time processing requirements.
- Collaborate with research scientists to deploy state-of-the-art deep learning models into production environments.
- Implement robust monitoring, logging, and alerting systems to ensure high availability and reliability.
- Drive the technical roadmap for our AI infrastructure, anticipating needs for 2026 and beyond.
- Conduct code reviews and mentor junior engineers to foster a culture of technical excellence.
Qualifications
- Master's degree in Computer Science, Engineering, or a related field (or equivalent practical experience).
- 5+ years of experience in software engineering, with a focus on machine learning operations (MLOps) or DevOps.
- Strong proficiency in Python, PyTorch, TensorFlow, or JAX.
- Deep understanding of distributed systems, containerization, and cloud platforms (AWS, GCP, or Azure).
- Experience with database technologies such as PostgreSQL, Elasticsearch, or time-series databases.
- Excellent problem-solving skills and the ability to work effectively in cross-functional teams.