Site Reliability Engineer (SRE)
- Джерело:
- djinni.co
Що робити
- Production Reliability – Help maintain the availability, stability, and performance of our production platform while proactively identifying opportunities to improve system reliability.
- Monitoring & Observability – Improve monitoring, dashboards, and alert quality to increase production visibility and reduce alert fatigue.
- Automation – Build scripts and automation solutions that reduce manual operational work, improve engineering efficiency, and enhance production reliability.
- Production Operations & Incident Response – Participate in troubleshooting production issues, perform root cause analysis, and contribute to long-term reliability improvements.
- Cloud Platform Operations – Support and improve our Kubernetes-based cloud platform, CI/CD pipelines, and production infrastructure running on GCP and AWS.
Що очікуємо
- 2–3 years of experience in Site Reliability Engineering, DevOps, Cloud Operations, Platform Engineering, or a similar role.
- Hands-on experience with Kubernetes and containerized environments.
- Familiarity with public cloud platforms (GCP or AWS).
- Experience with Linux systems and basic networking concepts.
- Experience with scripting or programming (Python, Bash, or similar).
Що пропонуємо
- Experience with CI/CD pipelines.
- Familiarity with Infrastructure as Code tools such as Terraform or Ansible.
- Experience with messaging technologies such as Kafka, Pub/Sub, or Redis.
- Exposure to Canary, Blue/Green, or Feature Flag deployment strategies.
- Understanding of Site Reliability Engineering principles, including SLIs, SLOs, and error budgets.
Схожі вакансії
З блогу Trackr
Усі статті →Знайдено через trackr.help/jobs · Канал: @trackrhelp · Бот для персональних сповіщень: @trackrhelpBot


