Job Description
We are defining the infrastructure for the next decade. Nexus Horizon Systems is seeking a visionary Senior AI Infrastructure Architect to lead our breakthrough projects in autonomous systems and generative AI.
In this pivotal role, you will bridge the gap between cutting-edge algorithmic research and scalable production environments. We are building the '2026' standard of computing—where latency is near-zero, intelligence is ubiquitous, and systems are self-healing. If you are passionate about the future of technology and possess a deep understanding of large-scale machine learning systems, we want to hear from you.
Responsibilities
- Architect Scalable AI Systems: Design and implement robust infrastructure for training and deploying large-scale generative models and autonomous agents.
- MLOps Strategy: Establish end-to-end MLOps pipelines, ensuring seamless model lifecycle management from experimentation to production.
- Performance Optimization: Drive initiatives to reduce inference latency and optimize resource utilization across distributed clusters.
- Cloud Native Leadership: Leverage Kubernetes, Docker, and serverless architectures to build resilient, fault-tolerant systems.
- R&D Collaboration: Partner with research scientists to translate theoretical models into high-performance, real-world applications.
Qualifications
- Education: Master’s or PhD in Computer Science, Mathematics, or a related technical field, or equivalent practical experience.
- Core Expertise: Extensive experience with deep learning frameworks (PyTorch, TensorFlow) and high-performance computing.
- Infrastructure: Strong proficiency in Linux, containerization (Docker, Kubernetes), and cloud platforms (AWS, GCP, or Azure).
- Programming: Proficiency in Python, Rust, or C++ for performance-critical applications.
- Problem Solving: Demonstrated ability to architect solutions for complex, unstructured problems in distributed systems.