Site Reliability Engineer ID53670
hace 3 días
Riba-roja de Túria
ppAgileEngine is an Inc. 5000 company that creates award-winning software for Fortune 500 brands and trailblazing startups across 17+ industries. We rank among the leaders in areas like application development and AI/ML, and our people-first culture has earned us multiple Best Place to Work awards. /p h3WHY JOIN US /h3 pIf you're looking for a place to grow, make an impact, and work with people who care, we'd love to meet you! /p h3ABOUT THE ROLE /h3 pWe are looking for a Middle SRE Operations Engineer to maintain reliability across a cloud-based SaaS platform. You’ll handle live incidents, improve observability, and reduce toil through automation using Kubernetes, Terraform, Grafana, and AWS. Hands‑on, execution‑focused, with real ownership across CI/CD pipelines, GitOps workflows, and on‑call rotations. /p h3WHAT YOU WILL DO /h3 ul liMonitor and support production and staging environments to ensure availability, performance, and stability. /li liRespond to incidents, perform triage and root cause analysis, and contribute to remediation efforts. /li liParticipate in on-call rotations with defined SLAs. /li liHandle operational requests from internal teams. /li liMaintain and improve monitoring, alerting, dashboards, logs, and metrics. /li liSupport CI/CD pipelines, production releases, and GitOps workflows. /li liContribute to automation initiatives to reduce operational overhead. /li liMaintain and improve Kubernetes‑based infrastructure and containerized workloads. /li liSupport Infrastructure as Code practices and environment improvements. /li /ul h3MUST HAVES /h3 ul lib2+ years of experience /b in Site Reliability Engineering, DevOps, or Production Operations. /li liExperience with bAWS /b supporting production environments. /li liExperience supporting bproduction SaaS applications /b. /li liStrong understanding of bCI/CD systems /b (GitHub Actions, Jenkins, CircleCI). /li liExperience with bGitOps and Git fundamentals /b. /li liExperience using bGitHub, Jira, and Confluence /b. /li liExperience with bKubernetes /b (EKS, kOps or similar). /li liExperience with bDocker and containerization /b. /li liExperience with bobservability tools /b (Grafana, Prometheus, Loki, PagerDuty). /li liProficiency in bscripting /b (Bash, Python, or Go). /li liExperience with bInfrastructure as Code /b (Terraform, Helm). /li liAbility to work within structured operational processes and SLAs. /li liStrong written and verbal English communication skills. /li liSelf‑driven with a growth mindset. /li /ul h3NICE TO HAVES /h3 ul liAWS certifications such as Solutions Architect, DevOps Engineer, or SysOps Administrator. /li liExperience with multi‑tenant SaaS environments. /li liExperience working in globally distributed teams. /li liFamiliarity with ChatOps practices. /li liExperience improving monitoring quality and reducing alert fatigue. /li /ul h3PERKS AND BENEFITS /h3 ul libProfessional growth /b: Mentorship, TechTalks, and personalized growth roadmaps. /li libCompetitive compensation /b: USD‑based pay with education, fitness, and team activity budgets. /li libExciting projects /b: Modern solutions with Fortune 500 and top product companies. /li libFlextime /b: Flexible schedule with remote and office options. /li /ul /p #J-18808-Ljbffr