Company Description
An AI technology company developing enterprise-grade artificial intelligence solutions that help organizations automate workflows, improve decision-making, and accelerate digital transformation. Its platforms combine generative AI, machine learning, and data intelligence to solve complex business challenges across industries.
- Location Chennai
- Industry Technology / IT
- Experience Range 7+ years
- Must-have Skills AWS / Azure / GCP, Generative AI, Hugging Face Transformers, LangChain / LlamaIndex, Large Language Models (LLMs), MLOps, OpenAI / Anthropic / Mistral / LLaMA, Prompt Engineering, Python, PyTorch / TensorFlow, Retrieval-Augmented Generation (RAG), Vector Databases (FAISS / Pinecone / Weaviate / Milvus)
Job Summary
The Senior Generative AI Engineer is responsible for designing, developing, and deploying enterprise-grade Generative AI solutions using Large Language Models (LLMs), Retrieval-Augmented Generation (RAG), multimodal AI, and modern AI orchestration frameworks. The role focuses on building scalable GenAI pipelines, optimizing AI model performance, integrating production-ready AI solutions, and collaborating with cross-functional teams to deliver innovative AI-powered applications.
Key Responsibilities
- Develop and fine-tune Generative AI models including GPT, LLaMA, Claude, and Stable
- Diffusion for text, image, audio, and video generation.
- Design and build end-to-end Generative AI pipelines covering data collection, preprocessing, model training, evaluation, and deployment.
- Integrate Large Language Models (LLMs) and multimodal AI models into production systems using LangChain, LlamaIndex, OpenAI APIs, or similar frameworks.
- Design and implement Retrieval-Augmented Generation (RAG) solutions for enterprise use cases.
- Collaborate with Data Science, Product, and Engineering teams to build AI-driven applications.
- Research and evaluate emerging AI technologies, tools, and frameworks to drive innovation.
- Optimize model inference, response latency, scalability, and cost efficiency for production deployments.
- Support deployment and operationalization of enterprise AI solutions while following MLOps best practices.
Required Skills & Experience
- Bachelor’s or Master’s degree in Computer Science, Artificial Intelligence, Machine Learning, Data Science, or a related field.
- 7+ years of experience in Machine Learning/Deep Learning Engineering with at least 5+ years of experience in Generative AI applications.
- Strong proficiency in Python programming.
- Hands-on experience with Machine Learning frameworks such as PyTorch or TensorFlow.
- Strong experience with Large Language Models (LLMs) including OpenAI, Anthropic, Mistral, and LLaMA.
- Experience using LangChain, Hugging Face Transformers, LlamaIndex, or similar AI frameworks.
- Strong understanding of Prompt Engineering techniques.
- Experience working with Embeddings and Vector Databases such as FAISS, Pinecone, Weaviate, or Milvus.
- Familiarity with cloud platforms including AWS, Azure, or Google Cloud Platform (GCP).
Knowledge of MLOps best practices. - Strong analytical, research, debugging, and problem-solving skills.
Preferred Qualifications
- Experience with diffusion and image generation models such as Stable Diffusion, Midjourney, or DALL·E.
- Exposure to multimodal AI systems combining text, vision, and speech.
- Experience with Azure AI Foundry and Azure AI Search.
- Knowledge of API development, Docker, Kubernetes, and CI/CD pipelines.
- Contributions to open-source AI projects or research publications in NLP or Generative AI.
Other Requirements
- Excellent communication and interpersonal skills.
- Self-motivated and capable of working collaboratively within distributed teams.
- Ability to coordinate effectively with onsite and offshore teams.
- Strong problem-solving mindset with a proactive approach toward technical challenges.