Senior Data Software Engineer/ Databricks, Apache Spark, PySpark

EPAM Systems Inc
  • Atlanta, GA
    18 days ago

    Job Description

    Back to Search

    Senior Data Software Engineer/ Databricks, Apache Spark, PySpark

    Remote in Georgia, & 4 others

    Data Software Engineering

    apply

    FacebookLinkedInSend via email

    Looking for something else?

    Find a vacancy that works for you. Send us your CV to receive a personalized offer.

    Find me a job

    Location-specific conditions & benefits*

    Choose an option

    We are seeking a skilled Senior Data Software Engineer with strong expertise in PySpark, SQL, and unit testing to join our data engineering team. The ideal candidate will have hands-on experience with Apache Spark, preferably within the Databricks environment, and will be responsible for building scalable data pipelines, optimizing data workflows, and ensuring code quality through rigorous testing practices.

    Responsibilities

    • Design, develop, and maintain scalable data pipelines using Apache Spark (Databricks preferred)

    • Write efficient and optimized PySpark code for data transformation and processing

    • Develop and execute complex SQL queries for data extraction, validation, and reporting

    • Implement unit tests using pytest to ensure code reliability and maintainability

    • Collaborate with data scientists, analysts, and other engineers to deliver high-quality data solutions

    • Monitor and troubleshoot data workflows and performance issues

    • Document technical designs, processes, and best practices

    Requirements

    • 3+ years of experience in Data Software Engineering

    • Proven experience with Apache Spark, ideally in a Databricks environment

    • Proficiency in PySpark and SQL

    • Background in unit testing frameworks, especially pytest

    • Understanding of data engineering principles and ETL processes

    • Familiarity with version control systems (e.g., Git)

    • Ability to work independently and in a collaborative team setting

    • Excellent problem-solving and communication skills

    • Proficiency in English at an Upper-Intermediate level (B2) or higher

    Nice to have

    • Experience with cloud platforms (e.g., Azure, AWS, GCP)

    • Knowledge of CI/CD pipelines and DevOps practices

    • Familiarity with Delta Lake, MLflow, or other Databricks-native tools

    Numbers & Facts

    LocationAtlanta, GA

    Skills

    • Apache Sparkunmatched
    • Best Practicesunmatched
    • Communication Skillsunmatched
    • Continuous Deployment/Deliveryunmatched
    • Continuous Integrationunmatched
    • Data Analysisunmatched
    • Data Managementunmatched
    • Data Qualityunmatched
    • Data Scienceunmatched
    • Database Extract Transform and Load (ETL)unmatched
    • DevOpsunmatched
    • English Languageunmatched
    • Gitunmatched
    • Identify Issuesunmatched
    • Problem Solving Skillsunmatched
    • Pytestunmatched
    • Quality Assurance Methodologyunmatched
    • SQL (Structured Query Language)unmatched
    • Scalable System Developmentunmatched
    • Software Engineeringunmatched
    • Source Code/Configuration Management (SCM)unmatched
    • Team Playerunmatched
    • Technical Writingunmatched
    • Technical/Engineering Designunmatched
    • Unit Testunmatched

    Be found by employers

    5,500+ employers search our resume database daily. Add yours to get found by recruiters looking for candidates like you.

    Level up your application

    Professional resume templates

    Browse dozens of recruiter approved resume templates, layouts and formats. Choose your favorite and make it your own in minutes.

    Free resume templates

    Free resume builder

    Improve your existing resume or start from scratch and create a standout, ATS-friendly resume. Add job-specific content, download and apply.

    Free resume builder