Systems SRE Data Center Engineer
Boingo · Frisco, TX · United States · On-site
Posted Oct 7, 2026
Sign up free: we match you to jobs like this, tailor your application and fill the form. 2 free applications every day.
Systems SRE Datacenter Engineer
The Systems SRE Datacenter Engineer is responsible for the daily operational maintenance, deployment, and reliability of Boingo’s global production, Staging, QA, Beta, and Dev environments. This hands-on role bridges physical datacenter operations, including racking, stacking, and cabling—with automated cloud-native systems administration.
Working as an individual contributor within an SRE and engineering team, you will support HPE Private Cloud Business Edition (PCBE) management platforms, maintain HPE VME and VMware virtualized workloads, monitor high-performance HPE Compute and Storage systems and run automated workflows on Ubuntu Linux to reduce administrative TOIL.
Basic Qualifications
Experience: 3 to 5 years of hands-on experience in systems administration, datacenter engineering, or Site Reliability Engineering (SRE).
Operating Systems: Solid, working proficiency with Ubuntu OS, RHEL 7+, and standard Linux command-line, system admin tools. Familiarity with Linux networking, systemd, permissions, and package management.
Datacenter Physical Layer: Direct experience performing physical datacenter hardware maintenance, including precision racking, server stacking, structured ethernet and fiber cabling, and label management.
HPE & Virtualization Fundamentals: Hands-on experience operating within HPE Private Cloud Business Edition (PCBE) or HPE GreenLake consoles. Experience supporting virtualized workloads hosted on VMware vSphere or HPE VME environments.
Education: Bachelor’s degree preferred.
Technical Skills & Capabilities
HPE Compute & Storage: Operational experience deploying, maintaining, and monitoring HPE Compute (ProLiant, Synergy) and HPE Storage systems such as HPE SimpliVity hyperconverged nodes or HPE Nimble Storage arrays. Basic familiarity with HPE iLO remote management APIs.
Kubernetes & Containers: Working experience deploying, troubleshooting, and maintaining containerized applications running on Kubernetes or RKE clusters.
Infrastructure as Code (IaC): Practical ability to execute, maintain, and update Ansible playbooks and Terraform scripts for routine server provisioning and configuration management.
Scripting: Proficiency in Python or Bash to write operational scripts, process logs, and automate repetitive system tasks.
Observability & Alerting: Hands-on experience using monitoring tools (Prometheus, Grafana, or InfluxDB) to respond to system alerts, track performance metrics, and troubleshoot outages.
Core Networking and Services: Basic understanding of DNS, LDAP, RADIUS, and network routing fundamentals in enterprise environments.
Agile Tooling: Daily familiarity with the Atlassian suite including Jira and Confluence
Duties & Responsibilities
Physical Datacenter Execution: Execute hardware rack-and-stack installations, structured cabling runs, power connection hookups, and component replacements (RAM, drives, NICs) in local and colocation…