Data Engineer

SPAR Information Systems LLC

  • Sunnyvale, CA
  • 8 days ago
  • Instant Apply
Want to know if you’re a fit?
Upload your resume and let our AI show you.

Skills

  • Amazon Simple Storage Service (S3)unmatched
  • Amazon Web Services (AWS)unmatched
  • Apacheunmatched
  • Apache Sparkunmatched
  • Apple Hardwareunmatched
  • Artificial Intelligence (AI)unmatched
  • Computer Systemsunmatched
  • Data Managementunmatched
  • Data Setsunmatched
  • Database Extract Transform and Load (ETL)unmatched
  • Distributed Computingunmatched
  • Hardware Designunmatched
  • Hardware Quality Assuranceunmatched
  • Identify Issuesunmatched
  • Problem Solving Skillsunmatched
  • Reliability Engineeringunmatched
  • SQL (Structured Query Language)unmatched
  • Scalable System Developmentunmatched
  • Startupunmatched
  • Systems Engineeringunmatched
  • Telecommunicationsunmatched
  • Validation Testingunmatched
  • Wireless Communicationsunmatched

Description

Role: Data Engineer Cellular Big data role

Location: Sunnyvale, CA/San Diego, CA (hybrid 3 days a week onsite)

Duration: 12+ Months

What the Team Does:

The Cellular Systems Engineering team captures and processes enormous volumes of log data generated from Apple devices during hardware validation and testing.
Historically only "bad" logs (known failures) were stored due to storage limitations.

Their new platform now stores both "good" and "bad" data at massive scale so AI/ML teams can identify hidden reliability issues and improve future silicon and cellular hardware designs.

Primary Responsibilities:
* Design and develop large-scale data pipelines
* Build scalable ETL/data ingestion solutions
* Curate and transform raw datasets
* Maintain production data pipelines
* Improve reliability and observability
* Design new infrastructure features
* Work directly with internal customers using the platform
* Build solutions while requirements continue to evolve
* Own projects end-to-end with minimal direction

Required Technical Skills

Must Have:

* Apache Spark
* Trino (Presto)
* Apache Airflow
* AWS
* SQL
* Large-scale Data Engineering
* Distributed Computing

* AWS Knowledge- Does NOT need to be an AWS expert.

Should understand:

* S3

* IAM Roles

* IAM Policies

* RBAC / Permissions

* Data governance

Type of Experience they're Looking For

This is NOT someone building small ETL jobs.

They're specifically looking for engineers who have experience processing:

* Terabytes of data per day
* Billions/trillions of records
* Distributed compute environments
* Production data infrastructure

Nice to Have:
* Cellular or wireless domain experience
* Telecommunications background
* Silicon or hardware data experience

These are bonuses-not requirements.

The ideal candidate:

* Comfortable with ambiguity
* Startup mentality
* Takes ownership
* Self-driven
* Strong problem solver
* Doesn't require handholding
* Comfortable designing solutions before requirements are finalized
* Moves quickly
* Enjoys building from scratch

Numbers & Facts

LocationSunnyvale, CA

Similar Jobs

See more jobs