Zorky CRMZorky CRM
EN|RU
@termdocs
← Все вакансии

Site Reliability and Observability Engineer

Central and Western District, Hong Kong, HKСкор 55/1001д назад
Аналитика рынка
📊 DevOps / SRE: зарплаты и спрос на рынке
Стек
site reliabilityclouddevopssre
Откликнуться
Загрузите резюме — мы свяжем вас с работодателем напрямую через нашу базу.
Отправить резюме →
Описание
Responsibilities: Build and maintain observability using logs, metrics, and dashboards for on-prem and cloud systems. Drive reliability improvements with performance analysis, capacity planning, stress tests, restore drills, failure prevention. Fine-tune alerts to reduce noise and improve signal quality and response speed; validate monitoring/alerts after releases and during deployments. Define and track reliability targets (SLA/RTO) for critical services. Improve incident response by maintaining operational procedures, service catalogs and clear escalation paths. g. health checks, remediation, validation gates. Perform Root Cause Analysis for reliability incidents and implement preventative actions. Ensure observability configuration changes are controlled and audit-evidenced. Uphold IT General Control and compliance standard with evidence retention, access controls, change approvals. Requirements Computer Science or related Engineering Degree (or above) 3 years or more working experience in SRE/DevOps duties
Контакты работодателя (email/phone/telegram) скрыты из публичного превью — отправьте резюме, чтобы мы связали вас напрямую.
Срочный вопрос? Напишите @termdocs