Site Reliability Engineer (Kubernetes) 4660

Tier4 GroupReston, United States
Full TimeOn-siteMidLimited info disclosed
37 views0 applications

Description

Site Reliability Engineer (Kubernetes) Location: Reston, VA (Hybrid – 2 days onsite) Type: Full-time We’re looking for a hands-on Site Reliability Engineer with a strong Kubernetes background to support and scale a modern containerized environment. This role focuses on maintaining reliable, high-performing Kubernetes platforms and supporting the tools and infrastructure around them. What You’ll Do Operate, maintain, and optimize Kubernetes/OpenShift environments Support containerized applications, networking, and cluster performance Build and improve automation using tools like Terraform and Ansible Enhance monitoring and observability (Prometheus, Grafana, Datadog) Partner with engineering teams to improve reliability, deployments, and CI/CD pipelines Support cloud infrastructure in Azure and hybrid environments What You Bring Strong hands-on Kubernetes experience (OpenShift is a plus) Experience with common Kubernetes ecosystem tools (monitoring, CI/CD, networking) Background in automation/IaC (Terraform, Ansible, or similar) Experience working in cloud environments (Azure preferred) Scripting skills (Bash, Python, or similar) Requirements: Minimum of 4-5 years of experience in a Kubernetes Engineering, Site Reliability Engineering, Platform Engineering or other similar role, with progressively increasing scope of responsibility. Extensive hands-on Kubernetes engineering experience, inclusive of managing workloads, deploying operators, configuring and managing routing/ingress, and overall cluster health and performance management. Expertise deploying and managing platform services and platform observability tools Show more Show less