۱ هفته پیش
استخدام Site Reliability Engineer (SRE) در شرکت اُکالا
حضوری
لیسانس
سابقه دارد (۱ سال)
حقوق توافقی
آقا و خانم
تمام وقت
مشاهده اطلاعات تماس
اطلاعات بیشتر
شرکت اُکالا در تهران جهت تکمیل کادر خود به افراد واجد شرایط ذیل نیازمند است.
| Description | job title |
| At Okala, we are looking for a Site Reliability Engineer (SRE) who will be responsible for ensuring system reliability, scalability, and operational excellence across our production environments. In this role, you will design and implement automation, improve infrastructure resilience, and maintain high availability of critical services. Your work will directly impact platform stability, deployment efficiency, and overall system performance. Responsibilities: Maintain and improve production system reliability, scalability, and availability Deploy updates, patches, and fixes while ensuring minimal service disruption Design and build automation tools to reduce manual effort and operational risk Perform root cause analysis for production incidents and implement preventive solutions Develop scripts to automate repetitive operational tasks Configure and maintain monitoring, logging, and alerting systems Design and maintain troubleshooting and preventive maintenance procedures Collaborate with cross-functional teams to resolve incidents efficiently Contribute to CI/CD pipeline improvements and deployment automation Participate in on-call rotations to support 24/7 operational environments Requirements: Bachelor’s degree in Computer Science, Engineering, or equivalent practical experience Minimum 1 year of hands-on experience in SRE, DevOps, or related roles Strong Linux system administration experience in production environments Proficiency in scripting and automation (Bash, Python, or similar) Solid understanding of databases, storage systems, and SQL Strong knowledge of containerization and orchestration (Kubernetes preferred) Experience designing and maintaining CI/CD pipelines (GitLab CI/CD preferred) Hands-on experience with monitoring, logging, and observability tools Experience with configuration management and infrastructure automation (e.g., Ansible) Strong understanding of networking fundamentals (DNS, routing, firewalls, load balancing) Experience with distributed systems and event streaming platforms (e.g., Kafka) Strong analytical and problem-solving skills Experience in incident response and reliability improvement Proactive and solution-oriented mindset Ability to collaborate effectively in cross-functional teams |
Site Reliability Engineer (SRE) |
متقاضیان واجد شرایط می توانند با کلیک روی لینک تکمیل فرم استخدام، رزومه خود را ارسال نمایند.
اطلاعات تماس
گزارش مشکل آگهی
- ثبتنام برای تکمیل فرم استخدام اینجا کلیک نمایید
- مهلت ۱۴۰۵/۰۶/۱۵
آگهیهای مشابه
جستجوهای مشابه
- استخدام مهندس IT در شهر تهران
- استخدام مهندس IT در استان تهران
- استخدام برنامه نویس در شهر تهران
- استخدام برنامه نویس در استان تهران
- استخدام مهندس کامپیوتر در شهر تهران
- استخدام مهندس نرم افزار در شهر تهران
- استخدام مهندس نرم افزار در استان تهران
- استخدام برنامه نویس پایتون (Python) در شهر تهران
- استخدام رشته فناوری اطلاعات (IT) در شهر تهران
- استخدام رشته فناوری اطلاعات (IT) در استان تهران
دستهبندی آگهیهای استخدام