Machine Learning Engineer - LLM & GenAI in India at Jobgether
Explore Related Opportunities
Job Description
This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Machine Learning Engineer - LLM & GenAI based in India.
Join an innovative engineering team building advanced AI solutions that transform how businesses interact with complex operational data. In this role, you will design and develop production-grade Large Language Model (LLM) applications that enable users to access insights through natural language conversations. You will work on cutting-edge Generative AI technologies, retrieval-augmented generation systems, and intelligent data workflows that create real-world impact. This global opportunity offers the chance to shape AI-powered products from architecture to deployment while collaborating with multidisciplinary teams. If you are passionate about machine learning, conversational AI, and building scalable intelligent systems, this role provides an exciting opportunity to contribute to the future of AI-driven analytics.
- Design, develop, and maintain LLM-powered backend services using Python and FastAPI to support conversational AI experiences.
- Build and optimize retrieval-augmented generation (RAG) pipelines that connect structured and unstructured data sources with intelligent language models.
- Implement frameworks such as LangChain, LlamaIndex, Haystack, or similar technologies to manage context retrieval, query routing, summarization, and response generation.
- Develop prompt engineering strategies, structured output generation workflows, and intent classification pipelines to improve AI performance and reliability.
- Integrate conversational AI capabilities with existing product interfaces and backend services while ensuring secure and efficient data flows.
- Design, test, and validate AI solutions across diverse analytics use cases, including data summaries, diagnostics, predictive insights, and performance analysis.
- Evaluate and improve retrieval accuracy, response quality, latency, and hallucination rates through continuous testing and optimization.
- Implement caching strategies, schema-based memory solutions, and automated evaluation pipelines to enhance system efficiency.
- Maintain technical documentation covering system architecture, APIs, prompt strategies, deployment processes, and model lifecycle improvements.
- 4–6 years of hands-on experience in Machine Learning, with proven experience building production-grade LLM or Generative AI applications.
- Strong proficiency in Python and experience developing backend services using FastAPI.
- Practical experience with LLM application frameworks such as LangChain, LlamaIndex, Haystack, or similar technologies.
- Experience designing retrieval pipelines and working with structured and unstructured data sources.
- Knowledge of SQL databases and data platforms such as PostgreSQL, CrateDB, or similar technologies.
- Understanding of vector databases including FAISS, Chroma, Pinecone, or comparable solutions.
- Ability to design effective prompting strategies, improve retrieval workflows, and optimize AI-generated responses.
- Experience evaluating AI systems through benchmarking, quality measurement, and performance optimization.
- Familiarity with OpenAI, Anthropic, Ollama, or other large language model platforms is preferred.
- Experience with embedding optimization, hallucination reduction techniques, AI orchestration frameworks, or multi-agent systems is a plus.
- Knowledge of EV analytics, fleet management, IoT data, or automotive technology is considered an advantage.
- Strong problem-solving skills, ownership mindset, and ability to work effectively in a remote global environment.
- Fully remote work opportunity with flexibility across locations.
- Opportunity to build cutting-edge AI and Generative AI solutions with real-world business impact.
- Direct influence on product architecture, technical decisions, and engineering practices.
- Work on challenging projects combining artificial intelligence, IoT, automation, and modern web technologies.
- Exposure to innovative applications in automotive technology and intelligent data analytics.
- Opportunity to collaborate with global teams and contribute to the growth of an emerging technology ecosystem.
- Continuous learning opportunities in rapidly evolving AI and machine learning fields.