Site Reliability Engineer (SRE)
- Location
- Not specified
- Employment
- Full-time
- Level
- Manager
- Category
- DevOps & SRE
- First seen
- Deadline
Description
We are hiring a Site Reliability Engineer to help define, measure, automate, and improve the uptime of critical customer-facing business processes.
This role is suitable for an engineer with strong SRE/DevOps experience. A software engineering background is not mandatory. The main focus is reliability from the customer and business perspective: service availability, transaction success, customer impact, operational continuity, and revenue protection.
You will work on SLA/SLO/SLI frameworks, business-flow monitoring, customer-impact alerting, automation, incident prevention, operational readiness, and reliability improvements together with development, product, observability, and platform teams.
The role includes participation in a structured on-call rotation for critical business services. The expectation is not only incident response, but also improving monitoring, automation, escalation, and preventive reliability controls after each case.
We are looking for a proactive engineer with strong ownership and the desire to introduce new tools, practices, dashboards, and automation that improve uptime and reduce production instability.
Apply at the source
This role was published by Bir and listed via Kapital Bank. Applications are handled there, not on this site.
Original posting: https://hr.kapitalbank.az/vacancies/8013