Skip to main content
EPAM Systems logo

Lead Data Software Engineer with Databricks with Apache Kafka, Apache Spark, Kubernetes

EPAM Systems
5 days ago
Remote
Georgia, Kazakhstan, Kyrgyzstan, Armenia, Uzbekistan, and United States

We are looking for a Lead Data Software Engineer to drive the migration of an existing data analytics platform, originally built on Azure resources such as Data Factory, Databricks, Event Hub, and Cassandra, to a cloud-agnostic platform deployable on-premises.

Responsibilities

  • Implement streaming and batch Spark pipelines running on Kubernetes
  • Lead the implementation of objects within the data generator framework
  • Oversee deployment and testing across local and development environments
  • Drive improvement of the current, actively changing solution
  • Conduct bug investigation and resolution
  • Perform unit testing to ensure solution quality
  • Participate in refinement, planning, and demo sessions
  • Guide and mentor team members throughout the implementation phase

Requirements

  • 5+ years of experience with Apache Spark, Confluent/Apache Kafka, and Kubernetes
  • Proficiency in Python
  • Familiarity with the data engineering domain and cloud-agnostic platform migrations
  • Capability to dive deeper into new technologies in a short time during active implementation phases
  • Ability to work independently with syncs with a team lead
  • Strong communication skills and proactiveness
  • Showcase of being a hands-on team player and technical leader
  • Proficiency in English at a B2+ level

Nice to have

  • Knowledge of Java
  • Familiarity with Kafka Connect, KSQL, and Kafka Streams
  • Skills in Ansible and Argo CD
  • Understanding of TDD
Job alerts

Not the right fit? Get jobs like this by email

New developer jobs, straight to your inbox. Instant alerts or a weekly digest.

Free. No spam, unsubscribe anytime.