GCB Services LLC logo

Platform Engineer

GCB Services LLC
  • Sunnyvale, CA
  • $70–$80 Per Year
  • Full-time
  • Employee
  • Instant Apply
1 day ago

Job Description

Role 1 - Core Platform Engineer (L1, breadth-first)

The engineering first line of defense: incident response, triage, reliability, and automation across the full infrastructure stack.

Day-to-day:

  • Runs incident response drills, post-mortems, and root cause analysis; learns from past incidents to prevent recurrence.
  • Starts the day reviewing overnight alerts and system performance metrics, triaging anomalies.
  • Participates in team stand-ups on projects, incidents, and daily priorities.
  • Automates routine processes, analyzes system logs, and builds tools to strengthen monitoring.
  • Works alongside software engineers advising on resilient-code best practices and reviewing changes pre-deployment.
  • Maintains high SLIs/SLOs; documents work and shares insights with a customer-centric mindset.

Must have:

  • Architecture, design patterns, reliability, and scaling of new and existing systems.
  • Incident command experience - driving RCA, coordinating cross-functional teams, ensuring corrective-action follow-through.
  • Observability built from the ground up - defining SLOs/SLIs, closing monitoring gaps, alerting strategies that catch failures before customers do.
  • Linux kernel internals - scheduler, memory allocation, driver subsystems.
  • High-quality code in at least one language (Python, Go, or similar).
  • System-level debugging - kdump, kernel panic analysis.
  • IaC (Ansible, Terraform, Kubernetes) and CI/CD (GitLab CI, AWX, etc.) for bare-metal or cloud infrastructure.
  • TCP/IP and network programming.
  • Distributed storage systems - object, block, and/or file storage paradigms.
  • Strong communication skills.

Nice to have:

  • Hardware and GPU troubleshooting.
  • OVN/OVS-based networking stack exposure.

Sourcing note: This is a deep SRE profile, not a pure generalist. The kernel-internals and system-level debugging bar is real and higher than a typical "L1" label implies - screen for genuine engineering depth, not helpdesk/NOC-tier breadth.

Role 2 - Platform Engineer (L2, depth-first)

Specialized domain expert embedded in a single foundation team: Storage, Compute, or SDN. Owns reliability, performance, scalability, and operational excellence for that one pillar. No cross-vertical work.

Must have:

  • Deep expertise in one domain (Storage / Compute / SDN).
  • SRE as a fundamental minimum - has set up observability, done alert management, and worked with SLIs/SLOs within that domain.
  • Ability to code the infrastructure: extend out-of-the-box product parameters to build observability layers, and tie SLAs/SLOs to business KPIs.
  • Domain specifics:

o Storage: block, blob, and file storage; distributed storage; performance diagnostics and data-path optimization.

Critical screening filter - operator vs. SRE:

Crusoe explicitly does not want another "operator" (storage admin doing patching, installs, upgrades). Screen hard for SRE substance (observability built, alerting, SLI/SLO ownership), not just domain tenure.

Role 3 - Platform Engineer (L2, depth-first)

Specialized domain expert embedded in a single foundation team: Storage, Compute, or SDN. Owns reliability, performance, scalability, and operational excellence for that one pillar. No cross-vertical work.

Must have:

  • Deep expertise in one domain (Storage / Compute / SDN).
  • SRE as a fundamental minimum - has set up observability, done alert management, and worked with SLIs/SLOs within that domain.
  • Ability to code the infrastructure: extend out-of-the-box product parameters to build observability layers, and tie SLAs/SLOs to business KPIs.
  • Domain specifics:

o Compute: Linux systems, KVM/QEMU, Cloud Hypervisor, kernel tuning, CPU/memory/VM optimization.

Critical screening filter - operator vs. SRE:

Crusoe explicitly does not want another "operator" (storage admin doing patching, installs, upgrades). Screen hard for SRE substance (observability built, alerting, SLI/SLO ownership), not just domain tenure.

Role 4 - Platform Engineer (L2, depth-first)

Specialized domain expert embedded in a single foundation team: Storage, Compute, or SDN. Owns reliability, performance, scalability, and operational excellence for that one pillar. No cross-vertical work.

Must have:

  • Deep expertise in one domain (Storage / Compute / SDN).
  • SRE as a fundamental minimum - has set up observability, done alert management, and worked with SLIs/SLOs within that domain.
  • Ability to code the infrastructure: extend out-of-the-box product parameters to build observability layers, and tie SLAs/SLOs to business KPIs.
  • Domain specifics:
  • SDN: OVS/OVN, network virtualization, high-performance networking, NIC tuning.

Critical screening filter - operator vs. SRE:

Crusoe explicitly does not want another "operator" (storage admin doing patching, installs, upgrades). Screen hard for SRE substance (observability built, alerting, SLI/SLO ownership), not just domain tenure.

Numbers & Facts

LocationSunnyvale, CA
Job TypeFull-time, Employee
IndustryTelecommunications Services
Salary$70–$80 Per Year
Company Size100 to 499 employees
HeadquartersSunnyvale, CA, US
Websitehttp://gcbservices.com/

About Company

GCB provides optimal engineering and business solutions for the wireless telecom industry by offering professional wireless network design, optimization and deployment services. Currently Wireless Industry is moving through a rapid metamorphosis. Availability and inventions of new technologies are constantly challenging wireless carriers and vendors to upgrade/ implement new systems and standards to better service the subscriber base.

GCB headquartered in Virginia is one of the premier independent providers of engineering & Information Technology services to the wireless telecom and IT industry. GCB was formed to provide the highest quality solutions while maintaining a fine balance between quality, cost, and timeline. The GCB team has extensive experience in wireless and Information Technology Services.

The GCB team has worked in all aspects of wireless projects spanning from initial planning, vendor and equipment evaluation, network dimensioning, network design, optimization, benchmarking, FCC compliance and overall project management.

At our IT house of services, we have been steadily progressing to achieve excellence in global IT Business Solutions and Consulting. Our consulting arm provides various Fortune 500 companies with integrated IT consulting services from concept to completion stage. Our business philosophy is founded on providing the best quality IT consulting, staff augmentation, quality assurance and training services to our valued clients. Whether you need helping hand in you Software Development Life Cycle (SDLC), Project Management, Human Resources Planning, or Employees Development and Training, we are always at arms length to help you with completing the tasks and addressing the challenges.

In essence, the GCB Team is qualified and capable to provide a broad set of the entire network and IT services for your business that range from business planning and capital budget modeling, to radio frequency engineering, system optimization, network system maintenance, data collection, and all aspects of integrated IT Services.

Skills

  • Analysis Skillsunmatched
  • Ansibleunmatched
  • Architectural Designunmatched
  • Automationunmatched
  • Best Practicesunmatched
  • CPU (Central Processing Unit)unmatched
  • Cloud Computingunmatched
  • Communication Skillsunmatched
  • Continuous Deployment/Deliveryunmatched
  • Continuous Integrationunmatched
  • Corrective Actionunmatched
  • Cross-Functionalunmatched
  • Customer/Client Researchunmatched
  • Debugging Skillsunmatched
  • Design Patterns Programming Methodologiesunmatched
  • Distributed Computingunmatched
  • Embedded Systemsunmatched
  • Follow Throughunmatched
  • GPU (Graphics Processing Unit)unmatched
  • Go Programming Language (Golang)unmatched
  • Hyperion Pillarunmatched
  • Hypervisorsunmatched
  • Identify Issuesunmatched
  • Incident Managementunmatched
  • Incident Responseunmatched
  • K Virtual Machine (KVM)unmatched
  • Kernel Programmingunmatched
  • Linux Kernelunmatched
  • Linux Operating Systemunmatched
  • Memory Hardwareunmatched
  • Memory Managementunmatched
  • National Intelligence Council (NIC)unmatched
  • Network Programmingunmatched
  • Performance Metricsunmatched
  • Python Programming/Scripting Languageunmatched
  • Root Cause Analysisunmatched
  • Schedule Developmentunmatched
  • Service Level Agreement (SLA)unmatched
  • Software Engineeringunmatched
  • Software Patchesunmatched
  • Systems Analysisunmatched
  • TCP/IP (Transmission Control Protocol/Internet Protocol)unmatched
  • Virtual Machine (VM)unmatched
  • Virtualizationunmatched

Be found by employers

5,500+ employers search our resume database daily. Add yours to get found by recruiters looking for candidates like you.

Level up your application

Professional resume templates

Browse dozens of recruiter approved resume templates, layouts and formats. Choose your favorite and make it your own in minutes.

Free resume templates

Free resume builder

Improve your existing resume or start from scratch and create a standout, ATS-friendly resume. Add job-specific content, download and apply.

Free resume builder