About this role
Utilize your monitoring and scripting skills as a DevOps Engineer in an L1 SRE role. Focus on maintaining system integrity through effective incident management and runbook applications.
We are hiring a DevOps Engineer specializing in SRE with 2–5 years of relevant experience. You will monitor system performance, conduct incident triage, and execute documented runbooks under the guidance of senior engineers. Your ability to communicate effectively during incidents will be critical to the team’s success in resolving issues promptly.
Key Responsibilities
Monitor system alerts using Prometheus and Grafana
Execute runbooks for handling incidents and maintenance
Provide initial incident triage and escalation support
Document all incidents and identify areas for process improvement
Assist in onboarding current applications to the framework
Requirements
2–5 years in IT operations or DevOps engineering
Extensive knowledge of networking and security protocols
Experience with Kubernetes management
Familiarity with documentation best practices
Strong problem-solving and analytical skills
Contribute to efficient operations in this pivotal role with a focus on continuous improvement and automation.