۲ ساعت پیش

استخدام Site Reliability Engineer | Production در اسنپ
اسنپ

استخدام Site Reliability Engineer | Production در اسنپ

اسنپ
تهران (زعفرانیه)
اطلاعات تماس

مختلف
مقطع تحصیلی اعلام نشده
سابقه دارد (۲ سال)
حقوق توافقی
آقا و خانم
تمام وقت

مشاهده اطلاعات تماس
اطلاعات بیشتر
امروز

اسنپ در تهران (زعفرانیه) جهت تکمیل کادر خود از واجدین شرایط زیر دعوت به همکاری می نماید:

Description job title
Ready to Get on Board?
Help us shape the future of ride-hailing and urban mobility. Submit your CV and let's build smarter cities together.
What You’ll Drive Forward
In this role, you will help scale and stabilize our systems as we grow. As part of the SRE team, you'll work on automating operations, managing incidents, and supporting the infrastructure that enables our developers and QA teams to build and release with confidence.
You'll be responsible for improving system reliability, monitoring, and observability while ensuring high availability across environments. This role includes participation in a 24/7 shift or on-call rotation.
Manage Incidents: Respond to incidents, perform root cause analysis, and help drive resolution and recovery.
Monitor & Alert: Improve and tune monitoring systems (Grafana, Prometheus) to ensure issues are detected early.
Participate in On-Call: Join a rotating on-call schedule to monitor systems and respond to critical alerts.
Collaborate Across Teams: Work closely with developers, QA, and product engineers to support releases and operational improvements.
Improve Stability: Proactively identify and fix reliability issues that could affect production uptime.
Automate Operations: Build scripts and tools to eliminate manual work and reduce operational overhead.
Deploy Services: Assist in deploying and maintaining services across staging and production environments.
Support Staging: Troubleshoot and resolve issues in pre-production environments to unblock QA and development teams.
What Powers Your Drive
At least 2 years of experience in a DevOps, SRE, or infrastructure engineering role.
Solid understanding of SRE principles: SLIs, SLOs, SLAs, Error Budgets.
Experience with Python (or another scripting language).
Hands-on experience with CI/CD tools and pipelines.
Comfortable with Linux systems administration.
Experience with monitoring and observability tools: Prometheus, Grafana.
Familiarity with logging stacks (e.g., ELK, Loki) and tracing tools (Jaeger, Tempo).
Knowledge of databases such as PostgreSQL/MySQL and Redis.
Practical experience with Kubernetes, Docker, and Helm.
Bonus Points:
Experience working in a microservices or distributed system environment.
Site Reliability Engineer | Production

متقاضیان واجد شرایط می توانند با کلیک روی لینک تکمیل فرم استخدام، رزومه خود را ارسال نمایند.

اطلاعات تماس
گزارش مشکل آگهی
https://iranestekhdam.ir/?p=3104557
ابتدای صفحه
مختصری درباره ایران استخدام

سایت ایران استخدام در تاریخ ۱۳۹۱/۱/۱۰ راه اندازی شد و با تلاش گروهی و روزانه مدیران و نویسندگان خود در جهت تبدیل شدن به مرجع بروز آگهی های استخدامی گام برداشت. سعی همیشگی همکاران ما ارائه مطلوب و با کیفیت آگهی های استخدامی خدمت بازدیدکنندگان محترم این سایت بوده است. ایران استخدام به صورت مستقل و خصوصی اداره می شود و وابسته به هیچ نهاد و یا سازمان دولتی نمی باشد، این سایت تنها منتشر کننده ی آگهی های استخدامی بوده و بنابراین لازم است که بازدید کنندگان محترم سایت خود نسبت به صحت و سقم اخبار منتشر شده در آن هوشیار باشند.

نماد اعتماد الکترونیکی
ارسال رزومه