Job Description
We are seeking a visionary Senior Generative AI Engineer to lead our research initiatives for the 2026 technology roadmap. Nexus Future Labs is at the forefront of next-generation intelligence, and we need a technical leader to architect scalable Large Language Models (LLMs) and drive the future of autonomous agents. If you are passionate about the intersection of deep learning and real-world application, we want to meet you.
In this pivotal role, you will define the architecture for our flagship AI products, mentor a team of brilliant engineers, and push the boundaries of what is possible in 2026 and beyond.
Responsibilities
- Architect and deploy state-of-the-art LLMs and Transformer-based models optimized for production environments.
- Design and implement Retrieval-Augmented Generation (RAG) pipelines to enhance model accuracy and context awareness.
- Collaborate closely with product managers and data scientists to translate business requirements into technical AI solutions.
- Optimize model inference latency and cost-efficiency using quantization and pruning techniques.
- Establish MLOps best practices for continuous training, evaluation, and deployment cycles.
- Conduct rigorous code reviews and technical mentoring for the engineering team.
Qualifications
- Bachelor’s or Master’s degree in Computer Science, Machine Learning, or a related field (PhD preferred).
- 5+ years of professional experience in machine learning, deep learning, or natural language processing.
- Expert proficiency in Python, PyTorch, or TensorFlow.
- Proven experience fine-tuning open-source models (e.g., Llama, Mistral, Falcon).
- Strong understanding of distributed systems and cloud infrastructure (AWS, GCP, or Azure).
- Excellent communication skills with the ability to explain complex technical concepts to non-technical stakeholders.