Skip to content
Columbus, OH
Search jobs, keywords, companies
Enter location or "remote"
SO Tracking Label
Search
Search
Sign up
Log in
Find Jobs
Salary Tools
Career Advice
Free Resume Templates
Free Resume Builder
Employers / Post Job
Find Jobs
Salary Tools
Career Advice
Resume
Free Resume Templates
Free Resume Builder
Employers / Post Job
See more jobs
X
Sr. Data Engineer | USC Only | W2 Contract | REMOTE.
Xlysi
Chicago, Illinois
3 days ago
Remote
Apply
Want to know if you’re a fit?
Upload your resume and let our AI show you.
Upload Resume
Have An Account? Log In
Skills
Amazon Simple Storage Service (S3)
unmatched
Amazon Web Services (AWS)
unmatched
Apache Spark
unmatched
Application Programming Interface (API)
unmatched
Artificial Intelligence (AI)
unmatched
Best Practices
unmatched
Cloud Computing
unmatched
Communication Skills
unmatched
Continuous Deployment/Delivery
unmatched
Continuous Integration
unmatched
Data Formats
unmatched
Data Lake
unmatched
Data Management
unmatched
Data Partitioning
unmatched
Data Processing
unmatched
Data Quality
unmatched
Data Science
unmatched
Data Storage
unmatched
Data Warehousing
unmatched
Database Design
unmatched
Database Extract Transform and Load (ETL)
unmatched
Database Technology
unmatched
Distributed Computing
unmatched
Distributed Databases
unmatched
Electronic Medical Records
unmatched
Git
unmatched
Linux Operating System
unmatched
Natural Language Parsing
unmatched
NoSQL
unmatched
Python Programming/Scripting Language
unmatched
REST (Representational State Transfer)
unmatched
Relational Databases (RDBMS)
unmatched
SNMP (Simple Network Management Protocol)
unmatched
SQL Databases
unmatched
Scala Programming Language
unmatched
Scalable System Development
unmatched
Source Code/Configuration Management (SCM)
unmatched
Team Player
unmatched
Telemetry
unmatched
Test Automation
unmatched
Test Data
unmatched
Unix Shell Programming
unmatched
Validation Testing
unmatched
+ show more
Description
Responsibilities:
Design and develop scalable ETL pipelines using Apache Spark (Scala/Python)
Build ingestion pipelines for network data (telemetry, syslogs, SNMP, configs, ticketing systems)
Transform raw data into query-ready formats in Data Lake
Implement monitoring, alerting, and pipeline reliability solutions
Develop CI/CD pipelines for data engineering workflows
Optimize large-scale data processing (batch & mini-batch, billions of events/day)
Manage data storage across distributed systems, databases, and APIs
Implement data quality checks, validation rules, and automated testing
Design and manage schemas, partitioning, and storage formats (Parquet, etc.)
Support data backfills and reprocessing for upstream or schema changes
Collaborate with data science and engineering teams for AI/ML data needs
Document data processes and ensure best practices across teams
Requirements:
Strong experience with Apache Spark (Scala or Python)
Hands-on experience building ETL pipelines at scale
Strong Python skills (Spark/PySpark preferred)
Experience with AWS (S3, Glue, Athena, EMR)
Strong SQL and relational database knowledge
Experience with Airflow or similar orchestration tools
Knowledge of data warehousing, partitioning, and columnar storage (Parquet)
Experience with data quality frameworks and validation
Linux and shell scripting experience
Experience with Git and version control
Strong communication and collaboration skills
Preferred:
Experience with Spark Streaming / Structured Streaming
Experience with Kafka or similar streaming platforms
NoSQL database experience
Experience with network data (syslogs, SNMP, telemetry)
REST API / cloud SDK integration (boto3, etc.)
Experience with automated testing for data pipelines
Knowledge of log parsing / text analytics
Telecom or large-scale network environment experienceS
Numbers & Facts
Location
Chicago, Illinois
(
Remote
)
Website
http://www.xlysi.com
Resume Resources
Free Resume Templates
Free Resume Builder
Similar Jobs