Site Reliability Engineer Jobs in Virginia
Site Reliability Engineer jobs in Virginia are consistently active, with strong demand across federal defense contracting, cloud services, and financial technology from entry-level engineers through staff and principal-level roles. The heaviest hiring concentrates in Northern Virginia, Richmond, and the Hampton Roads corridor, where employers like Booz Allen Hamilton, Leidos, and Capital One maintain large engineering organizations. The most sought-after specialties include Kubernetes orchestration, infrastructure-as-code, and observability engineering tied to cloud-native environments. Find a role that fits below and apply directly.
Find JobsOverview
Showing 5 of 45+ Site Reliability Engineer jobs











We believe that every experience is a memory that can last a lifetime. Experiences shape the way people feel about a company. And they greatly influence how likely people are to advocate, contribute, and stay. At Medallia, we are committed to creating a world where organizations are loved by their customers and their employees.
We empower exceptional people to create extraordinary experiences together.
Bring your whole self.
The Role and Team
As a Senior Site Reliability Engineer, you will play a key role in designing, operating, and evolving the platforms and services that power Medallia's global production environment. You will work across engineering teams to improve reliability, scalability, performance, and operational maturity while driving automation and platform improvements at scale.
This role is expected to provide technical leadership, influence engineering best practices, and help shape the future direction of our cloud-native infrastructure and operational strategy.
We are looking for engineers who think beyond day-to-day operations and continuously seek ways to increase engineering leverage.
Engineering Leverage
At Medallia, we believe great engineers amplify the impact of themselves and those around them.
Responsibilities:
- Design, build, and operate highly available, scalable, and secure production platforms.
- Partner with software engineering teams to improve application reliability, scalability, performance, and operational readiness.
- Lead complex incident investigations, root cause analyses, and reliability improvement initiatives.
- Design and implement automation, self-service capabilities, and platform solutions that reduce operational toil.
- Leverage AI-assisted engineering tools and automation platforms to accelerate troubleshooting, improve productivity, and reduce operational overhead.
- Identify opportunities to streamline operational processes through automation, AI-enabled workflows, and platform engineering practices.
- Drive adoption of SRE principles, reliability standards, and operational best practices across engineering organizations.
- Develop and maintain infrastructure-as-code, deployment automation, and operational tooling.
- Support and improve CI/CD and GitOps-based deployment workflows.
- Design observability strategies using monitoring, logging, tracing, and alerting platforms.
- Participate in architecture reviews and provide guidance on scalability, resiliency, and operational excellence.
- Mentor junior engineers and contribute to the technical growth of the broader engineering organization.
- Act as a force multiplier by creating reusable solutions, self-service capabilities, and engineering standards that increase the effectiveness of multiple teams.
- Drive adoption of AI-assisted engineering workflows and operational automation across the organization.
- Drive engineering leverage initiatives that improve the productivity, reliability, and effectiveness of multiple engineering teams.
- Influence the broader engineering organization through platform thinking, standardization, and operational simplification.
Qualifications:
- 5+ years of experience leading reliability, platform engineering, infrastructure, or cloud operations initiatives in production environments.
- Demonstrated experience operating and supporting large-scale production environments.
- Demonstrated experience with Kubernetes and containerized workloads in production environments.
- Demonstrated experience with cloud infrastructure platforms such as AWS, OCI, or GCP.
- Demonstrated Linux systems administration and troubleshooting skills.
- Demonstrated experience developing automation and tooling using Python, Go, Bash, or similar languages.
- Demonstrated experience with infrastructure-as-code technologies such as Terraform.
- Demonstrated experience designing and supporting CI/CD and GitOps workflows.
- Demonstrated understanding of networking fundamentals including DNS, load balancing, TLS/SSL, routing, and service networking.
- Demonstrated experience troubleshooting distributed systems and leading production incident response efforts.
- Demonstrated track record of reducing operational complexity through automation, platform engineering, or process transformation initiatives.
- Proven ability to influence technical decisions across teams and drive engineering improvements beyond direct ownership.
- Ability to participate in an on-call rotation supporting production systems.
- Experience with GitOps platforms such as ArgoCD.
- Experience operating multi-region or hybrid-cloud environments.
- Experience with observability platforms such as Prometheus, Grafana, Loki, OpenTelemetry, or similar technologies.
- Experience designing and operating platform engineering solutions and self-service infrastructure.
- Experience supporting high-scale SaaS environments.
- Understanding of release strategies such as canary, blue/green, progressive delivery, and feature flag-based deployments.
- Experience with capacity planning, performance engineering, and resilience testing.
- Familiarity with security, compliance, and regulatory requirements in production environments.
- Experience using AI-assisted development, automation, or operational tooling to improve engineering productivity and service reliability.
- Experience applying AI-assisted engineering workflows to improve productivity, reliability, or operational efficiency at scale.
- Experience designing platform engineering solutions that enable self-service and increase engineering leverage.
- Experience mentoring engineers and leading cross-functional technical initiatives.
- Demonstrated passion for automation, process improvement, operational excellence, and engineering scalability.
- Strong communication, collaboration, and stakeholder management skills.
Medallia is committed to equal pay and transparency. The annual base salary range for this position is $128,500 - $190,000. Please note that the salary range information provided is a general guideline and combines all of the distinct labor markets within the US. It is uncommon for an individual to be hired at or near the top of the range for their role and compensation decisions are dependent on a variety of factors. Medallia considers factors such as (but not limited to) scope and responsibilities of the position, candidate’s work experience, candidate’s work location, education/training, key skills, internal peer equity, external market data, as well as, market and business considerations when making compensation decisions.
Medallia also offers competitive health and wellness benefits, including but not limited to medical, dental, vision, 401(k), short-term and long-term disability, life and AD&D insurance, statutory leaves, paid parental leave, and paid holidays. Benefits and eligibility may vary by location and role.
At Medallia, we celebrate diversity and recognize the value it brings to our customers and employees. Medallia is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, age (40 and over), disability, genetic information, veteran status or military service, or any other status protected by state or local law. Individuals with a disability who need an accommodation to apply please contact us at ApplicantAccessibility@medallia.com. For information regarding how Medallia collects and uses personal information, please review our Privacy Policies. Applications will be accepted for 30 days from the date this role was posted or until the role has been filled.
See All 45 Site Reliability Engineer Jobs in Virginia
Find roles in Virginia that match your experience and apply in just a few clicks.
Find JobsSite Reliability Engineer Jobs by City in Virginia
Where Virginia roles are concentrated, by current openings.
Site Reliability Engineer Job Market in Virginia
A snapshot from current Virginia openings, updated as new roles post.
Who's Hiring



Top Industries Hiring
- Technology & Software
- Banking & Financial Services
- Retail
- Trucking
- Aerospace & Defense
What Virginia Employers Look For
The qualifications that appear most often in site reliability engineer jobs across Virginia.
- Bachelor's degree in computer science, systems engineering, or a related technical field
- Hands-on experience with Kubernetes, Terraform, or comparable infrastructure-as-code tooling
- Proficiency in at least one scripting or programming language such as Python, Go, or Bash
- Experience designing and maintaining CI/CD pipelines using tools like Jenkins, GitHub Actions, or ArgoCD
- Familiarity with cloud platforms such as AWS, Azure, or Google Cloud in production environments
- Active or obtainable security clearance, preferred or required by many Virginia defense and government contractors
Site Reliability Engineer Jobs in Virginia: Frequently Asked Questions
How do you become a site reliability engineer in Virginia?
Virginia has no state-issued license for site reliability engineers, so the path runs through education and credentials recognized by employers. Most hiring managers expect a bachelor's degree in computer science, information systems, or engineering. Industry certifications such as AWS Certified DevOps Engineer, Certified Kubernetes Administrator, or Google Professional Cloud DevOps Engineer strengthen any application. Northern Virginia's dense defense and cloud ecosystem makes obtaining a federal security clearance a meaningful differentiator for candidates early in their careers.
Which companies hire site reliability engineers in Virginia?
Employers hiring site reliability engineers in Virginia right now include Comcast, Medallia, and Umbra, based on current listings on Migrate Mate as of August 2026. Virginia's concentration of federal agencies, defense contractors, and cloud infrastructure operations creates steady, year-round demand for this role across a wide range of organization sizes.
Which Virginia cities have the most site reliability engineer jobs?
Reston, McLean, and Arlington have the most site reliability engineer openings in Virginia. Northern Virginia drives the largest share of postings because of its dense cluster of federal agencies, defense contractors, and hyperscale data center operations, while Richmond's growing fintech and insurance sectors and Hampton Roads' military and government IT presence round out the distribution.
Are there remote site reliability engineer jobs in Virginia?
Yes, and more than most fields. About 50% of site reliability engineer openings tied to Virginia are remote or hybrid as of August 2026, reflecting the fundamentally cloud-based and tool-mediated nature of the work. Roles focused on observability, incident response, and infrastructure automation tend to offer the most remote flexibility, while positions requiring access to classified networks or on-premises hardware typically require on-site presence.
How can I get hired as a site reliability engineer in Virginia with little or no experience?
The most realistic entry path is a junior DevOps or systems administrator role with one of Virginia's large defense contractors or cloud-forward employers, where internal mobility into SRE teams is common. Booz Allen Hamilton, SAIC, and General Dynamics IT all run associate and early-career engineering programs in Virginia that accept candidates without professional SRE experience. Building a public portfolio on GitHub demonstrating infrastructure-as-code projects and earning the Certified Kubernetes Administrator credential before applying gives candidates a concrete edge over peers with similar academic backgrounds.
Where can I find and apply to site reliability engineer jobs in Virginia?
You can find and apply to site reliability engineer jobs in Virginia on Migrate Mate, which lists current Virginia openings updated regularly. Search the listings, find roles that match your experience and preferred location, and apply directly to the ones that fit.
See All 45 Site Reliability Engineer Jobs in Virginia
Find roles in Virginia that match your experience and apply in just a few clicks.
Find Jobs