NE
Senior Java Site Reliability Engineer in McLean, VA
Published on CazVidat Nadeem Ehsan
Apply for Senior Java Site Reliability Engineer in McLean, VA. 16-20 years experience required. Contract role in Banking/Financial Services. Hybrid project.
Verified by CazVid3 months agoOpen until: Oct 20, 2026
Salary
$100 per hour
Location
McLean, Virginia, United States
Employment type
Contract
Workplace
Not provided
Senior Java Site Reliability Engineer in McLean, VA
$100 per hour
Job description
Role: Senior Java Site Reliability Engineer
Experience: 16-20 Years
Job Type: Contract
Project: Hybrid
Location: McLean, Virginia, United States
Industry: Banking / Financial Services
Key Responsibilities
- Support and maintain highly available production platforms across cloud and distributed environments. Drive incident management, root cause analysis, problem management, and platform stability initiatives.
- Monitor and maintain uptime of Java applications and microservices.
- Proactively identify and resolve application performance bottlenecks.
- Conduct root cause analysis (RCA) for application outages and incidents.
- Implement resiliency patterns including circuit breakers, retries, and failover mechanisms.
- Lead reliability engineering efforts focused on system availability, performance optimization, and operational excellence.
- Implement and enhance observability solutions including monitoring, logging, alerting, and incident response automation.
- Collaborate with development, infrastructure, and cloud engineering teams to improve deployment reliability and operational efficiency.
- Support infrastructure modernization, cloud transformation, and platform automation initiatives.
- Coordinate disaster recovery testing, resiliency validation, capacity planning, and production readiness reviews.
- Provide technical leadership and mentor offshore/onshore engineering teams.
Required Experience
- 16–20 years of experience in Site Reliability Engineering (SRE), Production Engineering, Platform Engineering, or Application Support.
- Strong experience supporting large-scale enterprise production environments with proven incident and problem management skills.
- Expertise in Java application support, microservices architecture, and cloud environments.
- Hands-on experience with resiliency patterns and observability tools.
- Proven leadership in reliability engineering and operational excellence initiatives.