Site Reliability Engineer @AXON-Networks
All Others
Salary usd 160,000 - 2..
Remote Location
Employment Type full-time
Posted 1mth ago

[Hiring] Site Reliability Engineer @AXON-Networks

1mth ago - AXON-Networks is hiring a remote Site Reliability Engineer. πŸ’Έ Salary: usd 160,000 - 200,000 per year πŸ“Location: USA, Canada

Role Description

The Site Reliability Engineer will improve the availability, performance, scalability and recoverability of AXON Networks cloud solutions. You will combine software engineering with hands-on NOC operations to make the complete cloud-to-device service path observable, supportable and resilient at fleet scale.

You will help establish practical SRE capabilities inside the NOC while partnering closely with Support, Operations, cloud and DevOps Engineering. You will participate in a sustainable on-call rotation and improve the NOC’s ability to diagnose customer-impacting issues.

Qualifications

  • 5+ years of experience in site reliability engineering, production engineering, DevOps, cloud infrastructure, systems engineering or a closely related role.
  • Strong software or automation skills in Python, Go, Java, Bash or a comparable language, with experience producing maintainable operational code.
  • Hands-on experience operating distributed production systems in a public cloud environment and troubleshooting across application, infrastructure, network and device-integration layers.
  • Experience with Google Cloud Platform, Oracle Cloud Infrastructure and production Kubernetes environments.
  • Experience with infrastructure as code and delivery tooling such as Terraform, Helm, Git-based CI/CD and policy-as-code.
  • Strong Linux, containers and Kubernetes fundamentals, including deployment behavior, resource management, networking and failure diagnosis.
  • Strong troubleshooting & debugging skills in Kubernetes platforms.
  • Experience with modern observability practices and tools across metrics, logs, traces, alerting, dashboards and synthetic monitoring.
  • Familiarity with Prometheus, Grafana, OpenTelemetry or equivalent observability ecosystems.
  • Familiarity with Apache Pulsar or similar distributed messaging and streaming platforms handling requests from millions of devices.
  • Experience participating in an on-call rotation and responding effectively to high-severity, customer-impacting production incidents.
  • Working knowledge of SLOs, error budgets, capacity planning, resilience engineering, change safety and blameless incident learning.
  • Strong networking knowledge, including TCP/IP, DNS, DHCP, TLS, routing, NAT, load balancing and systematic packet- or session-level troubleshooting.
  • Clear communication, disciplined documentation and the ability to collaborate across NOC, cloud, DevOps, firmware and service-provider teams.
  • Bachelor’s degree in computer science, engineering or equivalent practical experience.

Requirements

  • Experience supporting cloud-managed CPEs such as broadband gateways, routers, ONTs, Wi-Fi/mesh systems or similar edge devices in a service-provider environment.
  • Familiarity with TR-069/CWMP, TR-369/USP, TR-181 data models, ACS or USP controller platforms, device telemetry and remote lifecycle management.
  • Experience supporting messaging and streaming platforms such as Apache Pulsar or Kafka, APIs and highly available databases used in device-management control planes.
  • Understanding of access technologies such as GPON/XGS-PON, DOCSIS, Ethernet or fixed wireless and how CPE, ONTs and provider networks interact.
  • Experience with firmware rollout automation, canary or cohort deployments, fleet health analysis and safe rollback practices.
  • Experience building auto-remediation, safe self-service operations or internal reliability platforms.
  • Experience supporting multiple service-provider customers in a 24Γ—7 telecommunications, broadband or managed-network environment.

Benefits

This position is fully remote within North America. Please note that we are unable to offer visa sponsorship for this role.

Annual salary range: $160,000 - $200,000

Company Description

At AXON Networks, we promote equal opportunities in all our recruitment processes, ensuring non-discrimination on the basis of gender, age, origin, disability, or any other personal circumstances. We assess talent based on objective criteria and foster an inclusive and diverse working environment.

Join AXON Networks!

axon-networks.com

axon-networks.hire.trakstar.com

Before You Apply
️
remote Be aware of the location restriction for this remote position: USA, Canada
β€Ό Beware of scams! When applying for jobs, you should NEVER have to pay anything. Learn more.
Site Reliability Engineer @AXON-Networks
All Others
Salary usd 160,000 - 2..
Remote Location
Employment Type full-time
Posted 1mth ago
Apply for this position
Did not apply βœ“
Applied βœ“
Sent Follow-Up βœ“
Interview Scheduled βœ“
Interview Completed βœ“
Offer Accepted βœ“
Offer Declined βœ“
Application Denied βœ“
Unlock 125,000+ Remote Jobs
️
remote Be aware of the location restriction for this remote position: USA, Canada
β€Ό Beware of scams! When applying for jobs, you should NEVER have to pay anything. Learn more.
Apply for this position
Did not apply βœ“
Applied βœ“
Sent Follow-Up βœ“
Interview Scheduled βœ“
Interview Completed βœ“
Offer Accepted βœ“
Offer Declined βœ“
Application Denied βœ“
Unlock 125,000+ Remote Jobs
Γ—
Apply to the best remote jobs
before everyone else

Access 125,000+ vetted remote jobs and get daily alerts.

4.9 β˜…β˜…β˜…β˜…β˜… from 500+ reviews

⚑ 127,070+ remote jobs, refreshed hourly

πŸ”” Real-time alerts: Apply first, direct to employer

πŸ›‘οΈ Vetted companies, no scams, true remote only

Unlock All Jobs Now

Maybe later