DevOps & Site Reliability Engineer
VoltaGrid · Houston
Job description
About the role
We are looking for a DevOps & Site Reliability Engineer to design, build and maintain our cloud infrastructure and ensure the reliability of our services. The role works closely with engineering teams to create scalable, observable, and resilient systems while fostering a culture of operational excellence.
Key responsibilities
- Design, build, and maintain cloud infrastructure.
- Manage and optimise Kubernetes clusters and containerised workloads in production.
- Develop and maintain infrastructure‑as‑code using Terraform or equivalent tools.
- Build and improve CI/CD pipelines (GitHub Actions, GitLab CI, Jenkins, etc.) to enable fast, safe, and reliable deployments.
- Implement and maintain monitoring, alerting, and observability systems such as Prometheus, Grafana, or Datadog.
- Define and track SLIs/SLOs, participate in incident response, root‑cause analysis and blameless post‑mortems.
- Identify and eliminate toil through automation and self‑service tooling.
- Configure and maintain on‑prem bare‑metal servers, Linux‑based infrastructure and virtualised assets.
- Collaborate with development teams on system design, capacity planning and performance optimisation.
- Participate in on‑call rotations and ensure production readiness of new services.
Required profile
- 4+ years of experience in DevOps, SRE or infrastructure engineering roles.
- Strong experience with at least one major cloud provider (AWS preferred, GCP or Azure acceptable).
- Deep hands‑on experience with Kubernetes and Docker in production environments.
- Proficiency with infrastructure‑as‑code tools, particularly Terraform.
- Experience building and maintaining CI/CD pipelines (GitHub Actions, GitLab CI, Jenkins, etc.).
- Solid understanding of monitoring and observability (metrics, logs, traces).
- Strong scripting skills in Bash, Python or Go.
- Strong Linux systems administration skills (Ubuntu, RHEL/CentOS or similar).
- Experience with virtualization platforms, networking, DNS, load balancing and security fundamentals.
Required skills
- AWS
- GCP
- Azure
- Kubernetes
- Docker
- Terraform
- GitHub Actions
- GitLab CI
- Jenkins
- Prometheus
- Grafana
- Datadog
- Bash
- Python
- Go
- Ubuntu
- RHEL/CentOS
- Proxmox VE
- Virtualisation (VM provisioning, storage, networking)
- Networking, DNS, load balancing, security fundamentals
Questions fréquentes
Why are you reporting this job?
Explore further
Salaries, guides and searches in the United States.
Salary: DevOps Engineer Based on 23 job offers in the United StatesApply in 30 seconds
Enter your email to apply. An account will be created automatically.
By continuing, you accept our terms of use.
Already have an account? Login
A question about this job?
Ask it here: you will get the full job summary by e-mail, right away.
Published 1 month ago
Expires 6 days from now
54 views · 0 interested
Boost your chances
Upload your CV — we will match you with relevant openings.
Analyzing your CV...
VoltaGrid
Houston