Senior Site Reliability Engineer
Aviso de fuente externaen E-Solutions
Job Title: Senior Site Reliability Engineer (SRE)Location: Buenos Aires (CABA), Argentina – Hybrid (3 Days Onsite)Experience: 10+ YearsJob SummaryWe are seeking an experienced Senior Site...
Salario
No especificado
Ubicación
Buenos Aires, Argentina
Tipo de empleo
Tiempo completo
Modalidad
No especificado
Senior Site Reliability Engineer
Buenos Aires, Argentina
Descripción del empleo
Job Title: Senior Site Reliability Engineer (SRE)Location: Buenos Aires (CABA), Argentina – Hybrid (3 Days Onsite)Experience: 10+ YearsJob SummaryWe are seeking an experienced Senior Site Reliability Engineer (SRE) to support and enhance the reliability, scalability, and performance of enterprise applications and cloud infrastructure. The ideal candidate will have extensive experience in Site Reliability Engineering, Production Support, or Application Support, with strong expertise in AWS, automation, observability, and software development using Python or Java/Spring Boot. This role requires excellent communication skills and the ability to work collaboratively in a hybrid environment.Key ResponsibilitiesEnsure high availability, reliability, and performance of production applications and cloud infrastructure.Monitor, troubleshoot, and resolve complex production issues while minimizing downtime.Design, develop, and maintain automation tools and scripts to improve operational efficiency.Deploy, manage, and optimize cloud infrastructure on Amazon Web Services (AWS).Build and enhance observability solutions using monitoring, logging, and alerting tools.Develop and maintain applications, automation scripts, and services using Python or Java/Spring Boot.Perform root cause analysis (RCA) and implement preventive measures for recurring issues.Collaborate with development, infrastructure, and operations teams to improve system reliability.Participate in incident response, on-call support, and continuous service improvement initiatives.Create and maintain technical documentation, runbooks, and operational procedures.Required Skills10+ years of experience in Site Reliability Engineering (SRE), Production Support, or Application Support.Strong hands-on experience with Amazon Web Services (AWS).Experience in automation using Python or Java/Spring Boot.Strong knowledge of Linux systems and troubleshooting.Experience with observability, monitoring, logging, and alerting tools (e.g., Prometheus, Grafana, CloudWatch, Datadog, Splunk, ELK).Experience with scripting, automation, and infrastructure management.Strong analytical, troubleshooting, and problem-solving skills.Excellent verbal and written communication skills in English.Preferred QualificationsExperience with DevOps tools and CI/CD pipelines.Knowledge of Infrastructure as Code (Terraform or CloudFormation).Experience with containerization technologies such as Docker and Kubernetes.AWS certifications are a plus.Work ModelLocation: Buenos Aires (CABA), ArgentinaWork Type: Hybrid (3 days per week onsite)Language: Professional English (Required)
¿Es tuya esta vacante?
Reclámala gratis y recibe candidatos con video en CazVid.