We are seeking a highly skilled individual to help build and expand our next-generation data processing platform. This role will focus on designing and developing distributed data processing solutions using Apache Spark while leveraging an existing enterprise Kafka ecosystem. The successful candidate will play a critical role in scaling our data processing capabilities, enabling cloud adoption, and supporting strategic migration initiatives involving Databricks and AWS. This position offers the opportunity to work on high-volume, mission-critical data platforms processing large-scale real-time and batch workloads.
Required Skills & Qualifications
5 years of hands-on development experience with Apache Spark.
Strong expertise in Spark SQL, Structured Streaming, DataFrames/Datasets, performance tuning, and optimization.
Experience integrating Spark with Apache Kafka.
Strong development skills in Java and Python.
Experience designing and supporting large-scale distributed systems.
Solid understanding of software engineering principles, object-oriented design, and testing practices.
Experience working with relational and distributed data platforms.
Bachelor’s Degree
Prior work experience at client or in client's Industry
Applicants must be able to work directly for Artech on W2
Preferred Skills & Qualifications
Databricks experience (SQL).
AWS experience, including services such as S3, EMR, Glue, Lambda, ECS/EKS.