JobTarget Logo

GenAI/LLM Engineer in India at Jobgether

NewJob Function: Engineering
Jobgether
India, India
Posted on
New job! Apply early to increase your chances of getting hired.

Explore Related Opportunities

Job Description

GenAI/LLM Engineer

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a GenAI/LLM Engineer based in India.

This is an opportunity to build next-generation generative AI applications for enterprise search, summarization, content generation, and other language-driven use cases.
You will work at the intersection of natural language processing, large language models, and modern software engineering.
The role involves designing LLM application architectures, optimizing prompts, integrating embedding models, and developing production-ready AI solutions.
You will work with both commercial and open-source models while exploring fine-tuning and domain-specific adaptation techniques.
A strong focus on scalability, performance, cost efficiency, and response latency will be central to the role.
You will also collaborate across backend and AI engineering environments to integrate LLM capabilities into reliable software systems.
This position is well suited to an engineer who combines strong Python development skills with deep curiosity and practical expertise in generative AI.

Accountabilities:
  • Design and develop high-performance generative AI applications using commercial and open-source large language models, including models such as GPT-4, Claude, Llama 3, and Mistral.
  • Develop, test, and systematically refine prompt templates, system instructions, and few-shot learning examples to improve model accuracy, consistency, relevance, and compliance.
  • Prepare custom training datasets and implement parameter-efficient fine-tuning approaches such as PEFT and LoRA to adapt language models for domain-specific tasks and use cases.
  • Integrate LLM services into scalable backend and microservices architectures using efficient API frameworks, asynchronous processing, message queues, and supporting infrastructure.
  • Implement strategies such as semantic caching, token optimization, and model routing to improve production performance while managing inference costs and response latency.
  • Work with embedding models, vector indexes, and retrieval-oriented architectures to support robust enterprise search and other knowledge-intensive AI applications.
  • Establish systematic approaches for evaluating LLM outputs through automated evaluation suites, structured testing, and human-in-the-loop validation frameworks.
Requirements:
  • 3–7 years of relevant professional experience in software engineering, artificial intelligence, machine learning, or related technical fields, with strong hands-on exposure to generative AI and LLM applications.
  • Strong Python programming skills and extensive experience with LLM development frameworks such as LangChain, LlamaIndex, and/or Semantic Kernel.
  • Deep understanding of transformer architectures, tokenization methods, quantization techniques, and the execution and integration of open-source language models.
  • Practical experience building scalable REST APIs using technologies such as FastAPI or Flask and integrating AI services with message brokers, vector indexes, or other backend infrastructure.
  • Experience with LLM application architecture, prompt engineering, model adaptation, and production deployment of AI-powered applications.
  • Strong analytical and testing skills, with the ability to systematically evaluate model outputs using automated metrics, evaluation suites, and human validation approaches.
  • Bachelor’s degree or equivalent technical education, combined with strong problem-solving skills and the ability to work independently in a fast-moving AI engineering environment.
Benefits:
  • Annual compensation ranging from INR 2,400,000 to INR 3,800,000 CTC, depending on relevant experience, skills, and qualifications.
  • Full-time, mid-senior-level opportunity focused on cutting-edge generative AI and large language model technologies.
  • Remote work flexibility across India, with opportunities to work from Bangalore or Pune through hybrid or remote-flexible arrangements.
  • Opportunity to work with leading commercial and open-source LLMs and modern AI development frameworks.
  • Exposure to advanced areas including prompt engineering, fine-tuning, embeddings, model optimization, AI evaluation, and scalable LLM application architecture.
  • Opportunity to contribute to enterprise-focused AI products involving search, summarization, content generation, and other high-impact use cases.
How Jobgether works:
We use an AI-powered matching process to ensure your application is reviewed quickly, objectively, and fairly against the role's core requirements. Our system identifies the top-fitting candidates, and this shortlist is then shared directly with the hiring company. The final decision and next steps (interviews, assessments) are managed by their internal team.
We appreciate your interest and wish you the best!
Data Privacy Notice: By submitting your application, you acknowledge that Jobgether will process your personal data to evaluate your candidacy and share relevant information with the hiring employer. This processing is based on legitimate interest and pre-contractual measures under applicable data protection laws (including GDPR). You may exercise your rights (access, rectification, erasure, objection) at any time.
#LI-CL1

Job Location

India, India

Frequently asked questions about this position

Continue to apply
Enter your email to continue. You’ll be redirected to the employer’s application.
By clicking Continue, you understand and agree to JobTarget's Terms of Use and Privacy Policy.