Senior Systems Software Engineer, Observability and Telemetry Platform

Senior Systems Software Engineer, Observability and Telemetry Platform

Vollzeit Homeoffice möglich
J
  • Design, implement, and support operational and reliability aspects of a large-scale Observability & Telemetry collection platform, focusing on performance at scale, real-time monitoring, logging, and alerting
  • Engage in and improve the full lifecycle of services, from inception and design through deployment, operation, and refinement
  • Support services before launch through system design consulting, development of software tools, platforms, and frameworks, capacity management, and launch reviews
  • Maintain live services by measuring and monitoring availability, latency, and overall system health
  • Scale systems sustainably through automation and evolve systems by driving changes that improve reliability and velocity
  • Practice sustainable incident response and blameless postmortems
  • Participate in an on-call rotation to support production systems

Requirements

  • BS degree in Computer Science or a related technical field involving coding (e.g., physics or mathematics), or equivalent experience
  • 5+ years of experience with Infrastructure automation, distributed systems design, and designing and developing tools for running large-scale private or public cloud systems in production
  • 5+ years of experience delivering foundational infrastructure and observability platforms
  • Experience with one or more of: Python, Go, Perl, or Ruby
  • In-depth knowledge of Linux, networking, and containers
  • Experience in using or running large private and public cloud systems based on Kubernetes, OpenStack, and Docker
  • Experience running Grafana, OpenTelemetry, Prometheus, and similar observability-focused tools

Core Competencies

Demonstrates expertise in designing and implementing large-scale Observability and Telemetry platforms, with a strong focus on performance, reliability, and automation. Proficient in managing cloud systems and utilizing observability tools to ensure system health and operational excellence.

Highest-signal resume keywords

  • Infrastructure Automation
  • Distributed Systems Design
  • Python Programming
  • Kubernetes Management
  • Observability Tools Experience

Hard Skills

  • Infrastructure Automation
  • Distributed Systems Design
  • Python
  • Go
  • Perl
  • Ruby
  • Linux
  • Networking
  • Containers
  • Cloud Systems

Soft Skills

  • Incident Response
  • Collaboration

Certifications & Qualifications

  • BS Degree in Computer Science

Industry Keywords

  • Observability
  • Telemetry
  • Performance Monitoring
  • System Health
  • Automation

Tools & Technologies

  • Kubernetes
  • OpenStack
  • Docker
  • Grafana
  • OpenTelemetry
  • Prometheus

#J-18808-Ljbffr

Senior Systems Software Engineer, Observability and Telemetry Platform Arbeitgeber: Jobtailor

Als Front Office Supervisor in unserem dynamischen Team bieten wir Ihnen die Möglichkeit, in einem unterstützenden und freundlichen Arbeitsumfeld zu wachsen. Wir legen großen Wert auf die berufliche Entwicklung unserer Mitarbeiter und bieten regelmäßige Schulungen sowie die Chance, Verantwortung zu übernehmen. Unsere Lage ermöglicht es Ihnen, Teil einer lebendigen Gemeinschaft zu sein, während Sie gleichzeitig die Standards unseres Franchise-Partners einhalten und unseren Gästen einen unvergesslichen Aufenthalt bieten.

J

Kontaktdaten:

Jobtailor Recruiting-Team