PubMatic seeks Senior Software Engineers with expertise in AI agents and big data technologies including Hadoop, Spark, Scala, Kafka, and cloud solutions. Responsibilities include building scalable big data platforms, developing backend services with Java and AWS, maintaining data pipelines, designing GenAI-powered agents, integrating LLMs, managing GenAI workflows, collaborating with cross-functional teams, participating in Agile processes, supporting customers, and performing code reviews. Requirements include 1-5 years of Java/backend experience, strong CS fundamentals, big data tool experience, GenAI application expertise, ability to lead feature development, and a bachelor's degree in engineering or equivalent. The role is full-time, hybrid, located in Pune, India. Benefits include parental leave, healthcare, broadband reimbursement, and office amenities.
What you'll do
Build, design, and implement a highly scalable, fault-tolerant big data platform to process terabytes of data and provide customers with in-depth analytics
Develop backend services using Java, REST APIs, JDBC, and AWS
Similar jobs
More roles worth a look
Related opportunities based on specialty and working model so candidates can keep momentum.
Build and maintain Big Data pipelines using Spark, Hadoop, Kafka, and Snowflake
Architect and implement real-time data processing workflows and automation frameworks
Design and develop GenAI-powered agents for analytics, operations, and data enrichment using frameworks like LangChain, LlamaIndex, or custom orchestration systems
Integrate LLMs (OpenAI, Claude, Mistral) into existing services for query understanding, summarization, and decision support
Manage end-to-end GenAI workflows including prompt engineering, fine-tuning, vector embeddings, and retrieval-augmented generation (RAG)
Work closely with cross-functional teams to improve availability and scalability of large data platforms and PubMatic software functionality
Participate in Agile/Scrum processes such as sprint planning, retrospectives, backlog grooming, user story management, and work item prioritization
Collaborate with product managers on software features for the PubMatic Data Analytics platform
Support customer issues via email or JIRA, provide updates and patches
Perform code and design reviews
Requirements
1-5 plus years of coding experience in Java and backend development
Solid computer science fundamentals, including data structure and algorithm design, and creation of architectural specifications
Expertise in professional software engineering best practices for the full software development life cycle, including coding standards and code reviews
Hands-on experience with Big Data tools and systems like Scala Spark, Kafka, Hadoop, Snowflake
Proven expertise in building GenAI applications including LLM integration (OpenAI, Anthropic, Cohere, etc.), LangChain or similar agent orchestration libraries, prompt engineering, embedding, and retrieval-based generation (RAG)
Experience in developing and deploying scalable, production-grade AI or data systems
Ability to lead end-to-end feature development and debug distributed systems
Experience in developing and delivering large-scale big data pipelines, real-time systems and data warehouses preferred
Demonstrated ability to achieve stretch goals in an innovative and fast-paced environment
Demonstrated ability to learn new technologies quickly and independently
Excellent verbal and written communication skills, especially in technical communications
Strong interpersonal skills and a desire to work collaboratively
Bachelor’s degree in engineering or equivalent from a well-known institute/university