Job Description
Are you ready to architect the future of intelligence?
Vertex AI Innovations is on a mission to redefine human-machine interaction through cutting-edge Generative AI. We are seeking a visionary Senior Generative AI Engineer to lead the development of our next-generation Large Language Model (LLM) agents and autonomous systems. If you thrive in a fast-paced, high-impact environment and want to push the boundaries of what AI can achieve in 2026 and beyond, we want to hear from you.
Why Join Us?
- Work on state-of-the-art LLM architectures and fine-tuning strategies.
- Competitive equity package and top-tier benefits.
- Flexible remote-first culture with hubs in San Francisco and New York.
Your Role:
As a Senior Generative AI Engineer, you will be responsible for designing, training, and deploying scalable AI models that power our core product. You will bridge the gap between theoretical research and production-grade software engineering.
Responsibilities
- Design and implement robust Retrieval-Augmented Generation (RAG) pipelines to enhance model accuracy and reduce hallucinations.
- Optimize LLM inference latency and cost using techniques such as quantization, pruning, and model distillation.
- Collaborate with cross-functional teams (Product, Data Science, and Design) to translate complex requirements into technical specifications.
- Conduct rigorous A/B testing and evaluation of AI models to ensure performance benchmarks are met.
- Stay ahead of the curve by researching emerging AI methodologies and integrating them into our tech stack.
- Mentor junior engineers and contribute to the technical vision of the AI division.
Qualifications
- PhD or Masterβs degree in Computer Science, Machine Learning, or a related field.
- 5+ years of experience in software engineering, with at least 2 years specifically focused on AI/ML model development.
- Deep understanding of Transformer architectures, BERT, GPT, and other NLP models.
- Proficiency in Python, PyTorch, TensorFlow, and modern ML libraries.
- Experience with vector databases (e.g., Pinecone, Milvus, Weaviate) and message queues (e.g., Kafka, RabbitMQ).
- Strong grasp of distributed systems and cloud infrastructure (AWS, GCP, or Azure).