PubMatic seeks a senior engineer with 5-10 years experience in Generative AI and AI agent development. The role involves designing, developing, and deploying AI-driven features using Retrieval-Augmented Generation (RAG), vector databases (FAISS, Pinecone, Weaviate, Milvus), and large language models (LLMs). Responsibilities include leading technical design, implementing and optimizing LLMs, developing AI agents, prompt engineering, and collaborating with cross-functional teams. Required skills include strong knowledge of LLMs, AI agent architectures, agentic frameworks (LangGraph, CrewAI, AutoGen), vector databases, Python, TensorFlow, PyTorch, Hugging Face Transformers, and evaluation tools (Evals). A bachelor's degree in engineering is required. The position is full-time, hybrid (3 days in office, 2 days remote), based in Pune, Maharashtra, India. Benefits include parental leave, healthcare insurance, broadband reimbursement, and office amenities.
What you'll do
Lead design, development, and deployment of AI-driven features with end-to-end ownership
Similar jobs
More roles worth a look
Related opportunities based on specialty and working model so candidates can keep momentum.
Drive feasibility analysis, design specifications, execution, and release with quick iterations based on customer feedback in Agile environment
Spearhead technical design meetings and produce detailed design documents for scalable, secure, robust AI architectures
Align solutions with long-term product strategy and technical roadmaps
Implement and optimize LLMs including fine-tuning, deploying pre-trained models, and evaluating performance
Develop AI agents powered by RAG systems integrating external knowledge sources
Design, implement, and optimize vector databases for efficient and scalable vector search
Create and fine-tune sophisticated prompts to improve LLM performance
Utilize evaluation frameworks and metrics to assess and improve generative models and AI systems
Collaborate with data scientists, engineers, and product teams to integrate AI capabilities into products and tools
Stay updated with latest research and trends in LLMs, RAG, and generative AI technologies
Continuously monitor and optimize models for performance, scalability, and cost efficiency
Requirements
2 to 10 years of total experience with strong understanding of LLMs and their underlying principles — transformer architecture, attention mechanisms, and hyperparameter tuning
Proven experience designing and building AI agents, including multi-agent orchestration, tool-use patterns, multi-step planning, and agent memory architectures
Hands-on experience with agentic frameworks such as LangGraph, CrewAI, or AutoGen
Familiarity with RAG pipelines integrating external knowledge sources
In-depth knowledge of vector databases and indexing algorithms; practical experience with FAISS, Pinecone, Weaviate, or Milvus
Experience with agent observability, tracing, and guardrails using tools like Langfuse
Proficiency in prompt engineering for context-sensitive, domain-specific LLM outputs
Familiarity with Evals and other performance evaluation tools
Proficiency in Python and experience with machine learning libraries such as TensorFlow, PyTorch, and Hugging Face Transformers
Experience with data preprocessing, vectorization, and handling large-scale datasets
Ability to present complex technical ideas to technical and non-technical stakeholders
Bachelor’s degree in engineering or equivalent from a recognized institute
Tech stack
Generative AIAI agentsRetrieval-Augmented Generation (RAG)vector databaseslarge language models (LLMs)FAISSPineconeWeaviateMilvusLangGraphCrewAIAutoGenLangfusePythonTensorFlowPyTorchHugging Face TransformersEvalsDockerKubernetesAWSGCPAzure
Benefits
paternity/maternity leavehealthcare insurancebroadband reimbursementkitchen with healthy snacks and drinkscatered lunches
Apply now
Ready to take the next step in your career? Click the button below to continue to the application process.