Observability Engineer - Network Telemetry Specialist : 26-02084

Akraya Inc.

Redwood City, CA

JOB DETAILS
SALARY
$90–$95 Per Hour
SKILLS
Agile Programming Methodologies, Apache, Apache Kafka, Architectural Design, Architectural Services, Best Practices, Business Operations, Cloud Computing, Computer Science, Data Processing, Database Design, Distributed Computing, Emerging Technology, Engineering, High Availability, High Reliability, High Throughput, ITIL (IT Infrastructure Library), Incident Management, Java, Leadership, Mentoring, Network Administration/Management, Network Architecture/Engineering, Operations Processes, Performance Tuning/Optimization, PostgreSQL, Prototyping, Reporting Dashboards, Requirements Management, Root Cause Analysis, Scrum Project Management and Software Development, Software Development, Software Engineering, Standards Development, Technical Leadership, Technical Recruiting, Technical Strategy, Telemetry, Vehicle Driving
LOCATION
Redwood City, CA
POSTED
3 days ago
Primary Skills: Java (Expert), Apache Kafka (Expert), PostgreSQL (Expert), Grafana (Expert), Telemetry / Network Observability (Expert)
Contract Type: W2 Only
Duration: 12+ Months
Location: Redwood City, CA (Hybrid, remote id still an option with occasional travel)
Pay Range: $90 - $95 on W2

Job Summary:
We are seeking a Senior Staff Engineer – NPE Observability to lead the architecture and evolution of a large-scale telemetry and observability platform. This role will drive the technical strategy for real-time data ingestion, streaming, and visualization across a global network infrastructure. The ideal candidate will possess deep expertise in Java, Kafka, PostgreSQL, Grafana, and distributed systems, with a proven track record of designing highly scalable, low-latency observability platforms.
Key Responsibilities:
  • Architect and optimize scalable telemetry ingestion and storage platforms using Java and PostgreSQL.
  • Design and enhance high-throughput Apache Kafka streaming pipelines for real-time telemetry processing.
  • Define enterprise observability standards and build advanced Grafana dashboards for monitoring global infrastructure.
  • Architect stateful stream-processing solutions using technologies such as Apache Flink.
  • Evaluate and prototype emerging observability technologies including Model-Driven Telemetry (MDT), ClickHouse, and Thanos.
  • Define platform architecture, technical standards, and long-term engineering roadmaps.
  • Establish and monitor SLIs/SLOs to ensure high availability, reliability, and platform performance.
  • Lead complex root cause analysis, performance tuning, and architectural improvements for mission-critical systems.
  • Collaborate with software engineering, network engineering, and infrastructure teams to translate business requirements into scalable technical solutions.
  • Mentor engineering teams and drive technical excellence through architecture reviews and best practices.
Must-have Skills:
  • 10+ years of software engineering experience with expertise in distributed systems.
  • 5+ years of experience building large-scale network engineering, telemetry, or observability platforms.
  • Expert-level proficiency in Java backend development.
  • Strong expertise with Apache Kafka, including cluster architecture, messaging, and stream processing.
  • Advanced experience with PostgreSQL schema design, optimization, and performance tuning.
  • Expert-level experience developing enterprise dashboards using Grafana.
  • Strong understanding of distributed systems, real-time streaming, and high-throughput data processing.
  • Experience with Prometheus, Thanos, ClickHouse, or similar observability platforms.
  • Experience defining SLIs, SLOs, monitoring strategies, and incident management.
  • Strong stakeholder management, technical leadership, and architectural design skills.
Nice-to-have Skills:
  • Experience with Apache Flink or other stream-processing frameworks.
  • Knowledge of Model-Driven Telemetry (MDT) and modern telemetry architectures.
  • Experience working with cloud-native observability platforms and Kubernetes environments.
  • Familiarity with ITIL processes and enterprise operational best practices.
  • Experience mentoring senior engineers and leading technical strategy across large organizations.
Preferred Qualifications:
  • Bachelor's or Master's degree in Computer Science, Software Engineering, or a related technical discipline.
  • Experience designing globally distributed, high-availability observability platforms.
  • Strong background in Agile/Scrum software development methodologies.
  • Proven ability to drive long-term technical vision, innovation, and engineering excellence in enterprise-scale environments.
ABOUT AKRAYA
Akraya is an award-winning IT staffing firm consistently recognized for our commitment to excellence and a thriving work environmentMost recently, we were recognized Stevie Employer of the Year 2025, SIA Best Staffing Firm to work for 2025, Inc 5000 Best Workspaces in US (2025 & 2024) and Glassdoor's Best Places to Work (2023 & 2022)!

Industry Leaders in Tech Staffing
As Talent solutions provider for Fortune 100 Organizations, Akraya's industry recognitions solidify our leadership position in the IT staffing space. We don't just connect you with great jobs, we connect you with a workplace that inspires!

Join Akraya Today!
Let us lead you to your dream career and experience the Akraya difference. Browse our open positions and join our team!

About the Company

A

Akraya Inc.