Senior Staff Engineer

ICONMA, LLC
  • Redwood City, CA
  • $100.80–$105.80 Per Hour
  • Quick Apply
21 days ago

Job Description

Our client, a Digital Infrastructure company, is looking for a Senior Staff Engineer for their Redwood City, CA location.
 
Responsibilities:
  • The Senior Staff Engineer for NPE Observability is the preeminent technical strategist for Client’s global telemetry fabric. In this senior contract role, you will bridge the gap between high-scale distributed software and global network hardware, driving the architectural standards for our most complex data-intensive initiatives.
  • You will own the technical integrity of our streaming pipelines, ensuring telemetry from the global fleet is ingested, normalized, and processed with sub-second latency.
  • As a master of our tech stack (Java, Kafka, Postgres, Grafana), you will define the "Gold Standard" for technical excellence within the Network Platform Engineering (NPE) group.
Architectural Strategy & Technical Vision
  • Core Stack Evolution: Architect and optimize our primary ingestion and storage engines utilizing Java and PostgreSQL, ensuring high availability and performance at scale.
  • Real-Time Data Orchestration: Lead the design of high-throughput messaging systems using Apache Kafka to handle trillions of telemetry points with sub-second latency.
  • Unified Visibility: Define the global standard for observability visualization in Grafana, building complex, high-performance dashboards that aggregate data from diverse telemetry sources.
High-Scale Engineering & Innovation
  • Stream Processing Mastery: Architect massively parallel processing pipelines and stateful stream processing frameworks (utilizing tools like Apache Flink) to enable real-time anomaly detection.
  • Advanced R&D: Evaluate and prototype emerging technologies such as Model-Driven Telemetry (MDT) and ClickHouse/Thanos for long-term metric storage and high-cardinality data analysis.
  • Technical Roadmap Ownership: Drive the engineering team toward key milestones, ensuring the code we ship aligns with the 3–5 year long-term NPE vision.
Reliability & Systemic Leadership
  • Service Standards: Define and monitor critical SLI/SLO metrics (e.g., P95 response times) to ensure the platform maintains world-class performance and global ITIL compliance.
  • Incident Authority: Serve as the senior point of contact for complex root-cause analysis, identifying architectural weaknesses in the Java/Kafka/Postgres stack to prevent future outages.
  • Stakeholder Synthesis: Translate complex product requirements into deep technical specifications, managing relationships with both internal software teams and external network vendors.
 
Requirements:
  • Primary Skills: Java, PostgreSQL, Kafka, Kubernetes, SQL, Grafana.
  • Secondary Skills: Apache, Prometheus, Thanos, MDT (Model-Driven Telemetry).
  • 10+ years of professional experience in software engineering and distributed systems.
  • Domain Expertise: 5+ years of experience specifically in large-scale network engineering, telemetry, or observability platforms.
  • Java Expert: Mastery of Java for building high-performance, scalable backend services.
  • Data & Messaging: Deep expertise in PostgreSQL (schema design and tuning) and Apache Kafka (cluster architecture and stream management).
  • Visualization: Expert-level proficiency in Grafana for creating enterprise-level observability dashboards.
  • Large-Scale Systems: Proven experience with Prometheus, Thanos, or ClickHouse and working within a structured Agile/Scrum environment.
  • Bachelor’s or Master’s degree in Computer Science or a related technical field.
 
Why Should You Apply?

Numbers & Facts

LocationRedwood City, CA
Salary$100.80–$105.80 Per Hour

Skills

  • Agile Programming Methodologiesunmatched
  • Apacheunmatched
  • Apache Kafkaunmatched
  • Architectural Servicesunmatched
  • Computer Scienceunmatched
  • Data Analysisunmatched
  • Data Collectionunmatched
  • Database Designunmatched
  • Distributed Computingunmatched
  • Emerging Technologyunmatched
  • Health Planunmatched
  • High Availabilityunmatched
  • High Throughputunmatched
  • ITIL (IT Infrastructure Library)unmatched
  • Javaunmatched
  • Large-Scale Systemsunmatched
  • Leadershipunmatched
  • Messaging Technologyunmatched
  • Metricsunmatched
  • Network Architecture/Engineeringunmatched
  • Network System Hardwareunmatched
  • Parallel Computingunmatched
  • Performance Managementunmatched
  • PostgreSQLunmatched
  • Prototypingunmatched
  • Relationship Managementunmatched
  • Reporting Dashboardsunmatched
  • Research & Development (R&D)unmatched
  • Root Cause Analysisunmatched
  • Scalable System Developmentunmatched
  • Scrum Project Management and Software Developmentunmatched
  • Software Engineeringunmatched
  • Standards Developmentunmatched
  • Technical Strategyunmatched
  • Telemetryunmatched
  • Vehicle Fleetsunmatched

Be found by employers

5,500+ employers search our resume database daily. Add yours to get found by recruiters looking for candidates like you.

Level up your application

Professional resume templates

Browse dozens of recruiter approved resume templates, layouts and formats. Choose your favorite and make it your own in minutes.

Free resume templates

Free resume builder

Improve your existing resume or start from scratch and create a standout, ATS-friendly resume. Add job-specific content, download and apply.

Free resume builder