Job Description
Join 2026, a pioneering force in futuristic technology, to architect the infrastructure that powers tomorrow's AI. We are seeking a Senior AI Infrastructure Engineer to design scalable, high-performance systems for next-generation machine learning models. You will be at the forefront of computational efficiency, ensuring our solutions are robust, secure, and ready for the demands of the future.
What You Will Do:
- Architect and manage scalable cloud infrastructure for training and deploying massive language models.
- Optimize deep learning pipelines on GPU clusters to maximize throughput and minimize latency.
- Implement MLOps best practices, including CI/CD, automated model deployment, and monitoring.
- Collaborate with data scientists to translate research into production-grade software.
- Drive technical innovation in containerization, orchestration, and edge computing.
Responsibilities
- Design and maintain highly available distributed systems using Kubernetes and cloud-native technologies.
- Conduct performance tuning and resource allocation analysis for large-scale AI workloads.
- Develop scripts and tools to automate infrastructure provisioning and scaling.
- Ensure compliance with security standards and data privacy regulations.
- Mentor junior engineers and conduct code reviews to maintain high engineering standards.
Qualifications
- 5+ years of experience in software engineering, with a focus on backend systems or infrastructure.
- Strong proficiency in Python, Go, or C++.
- Deep experience with cloud platforms (AWS, GCP, or Azure) and serverless architectures.
- Expert knowledge of containerization (Docker) and orchestration (Kubernetes).
- Experience with high-performance computing (HPC) and GPU optimization (CUDA, cuDNN).
- BS in Computer Science, Engineering, or equivalent practical experience.