Full job description
Senior Software Engineer role in the CDP team at Epsilon. Responsible for delivering large-scale cloud-native data platforms on AWS using Databricks and distributed processing frameworks. Hands-on work with Python, Java, PySpark, Apache Spark, Databricks, AWS services, event-driven architectures, and SQL/NoSQL databases. Collaborate with global teams to align technical solutions with business objectives. Own full software development lifecycle including design, development, deployment, and documentation. Mentor junior engineers and participate in architecture and code reviews. Requires 5-8 years software engineering experience with expertise in data engineering, distributed systems, Databricks, Python, PySpark, Apache Spark, AWS, Kafka, SQL Server, MongoDB, CI/CD, DevOps, Terraform, and Ansible. Nice to have AWS and Databricks certifications and experience with Azure, GCP, and Generative AI technologies. Location: Wakefield, Massachusetts, USA. Salary range: $102,000 to $189,800 annually.
What you'll do
- Deliver large-scale cloud-native data platforms primarily on AWS leveraging Databricks and distributed processing frameworks
- Work hands-on across the technology stack including Python, Java, PySpark, Apache Spark, Databricks, AWS services, event-driven architectures, and SQL/NoSQL databases
- Partner with global engineering, product management, architecture, and business stakeholders
- Own the end-to-end software development lifecycle including requirements gathering, solution design, development, deployment, observability, and documentation
- Design and develop reusable, maintainable and scalable components
- Participate in architecture discussions, technical design reviews, and code reviews
- Mentor and guide junior engineers
- Foster a culture of innovation, accountability, collaboration, and technical excellence
Requirements
- B.E/B.Tech/M.Tech/MCA in Computer Science, Information Technology or related field
- 5-8 years of strong software engineering experience
- Expertise in large-scale data engineering and distributed systems architecture
- Experience in Data Warehousing, Data Lakes, Delta Lake architecture
- Hands-on expertise in Databricks, Python, PySpark, Apache Spark
- Experience with AWS services such as S3, Lambda, API Gateway, EventBridge
- Experience with Kafka and relational and NoSQL databases including SQL Server and MongoDB
- Experience implementing unit, integration, and regression testing
- Strong understanding of CI/CD and DevOps practices using Jenkins, GitHub/GitLab, Bitbucket
- Hands-on experience with Infrastructure as Code tools such as Terraform and Ansible
- Strong critical thinking and analytical skills
- Nice to have AWS and Databricks certifications
- Experience working with Azure and/or Google Cloud Platform
- Exposure to Generative AI technologies including LLMs, RAG architectures, and Agentic AI systems
Tech stack
PythonJavaPySparkApache SparkDatabricksAWSAWS S3AWS LambdaAPI GatewayEventBridgeKafkaSQL ServerMongoDBJenkinsGitHubGitLabBitbucketTerraformAnsibleCI/CDDevOpsData WarehousingData LakesDelta LakeDistributed SystemsCloud-nativeEvent-driven architecturesNoSQLGenerative AILLMsRAG architecturesAgentic AI
Benefits
Flexible time off (FTO)15 paid holidaysPaid sick timeParental/new child leaveChildcare & elder care assistanceAdoption assistanceComprehensive health coverage401(k)Tuition assistanceCommuter benefitsProfessional developmentEmployee recognitionCharitable donation matchingHealth coaching and counseling