Job Description
We are on a mission to define the technological landscape of 2026 and beyond. At Vertex AI Labs, we are building the next generation of generative intelligence systems. We are seeking a visionary Senior AI/ML Architect to lead our strategic roadmap and architectural initiatives.
In this role, you will not just maintain existing systems; you will architect the core infrastructure required for our products in the 2026 era. You will bridge the gap between theoretical research and production-grade engineering, ensuring our AI models are scalable, efficient, and future-proof.
Why Join Us?
- Shape the future of AI with cutting-edge technology.
- Competitive compensation package and equity options.
- Work with a world-class team of engineers and researchers.
- Focus on long-term impact and sustainable innovation.
Responsibilities
- Define the 2026 Roadmap: Lead the technical strategy for AI infrastructure upgrades, focusing on scalability and latency reduction for next-gen models.
- System Architecture: Design and implement robust, fault-tolerant distributed systems for training and deploying large-scale machine learning models.
- Model Optimization: Oversee the optimization of inference pipelines and fine-tuning strategies to maximize performance on edge devices and cloud environments.
- Tech Stack Leadership: Evaluate and integrate emerging technologies (e.g., neuromorphic computing, advanced LLMs) to maintain a competitive edge.
- Cross-Functional Collaboration: Partner with product managers and data scientists to translate business requirements into technical solutions.
- Code Review & Mentorship: Establish coding standards and mentor junior engineers to foster a culture of excellence.
Qualifications
- Education: Bachelor’s or Master’s degree in Computer Science, Mathematics, or a related field (PhD preferred).
- Experience: 7+ years of experience in software engineering, with at least 4 years focused on AI/ML architecture and deployment.
- Technical Skills: Deep proficiency in Python, PyTorch, TensorFlow, and modern GPU acceleration techniques.
- Cloud Expertise: Proven experience designing systems on AWS, GCP, or Azure.
- System Design: Strong understanding of distributed systems, microservices, and containerization (Kubernetes/Docker).
- Communication: Excellent verbal and written communication skills, with the ability to explain complex technical concepts to non-technical stakeholders.