We’re looking for a Mid to Senior DevOps / Site Reliability Engineer to design, build, and operate reliable, scalable, and secure cloud and/or on-prem infrastructure. You’ll work closely with engineering and security teams to automate delivery, improve system resilience, and embed security practices across the SDLC. Strong DevSecOps and security experience is a plus.
Key Responsibilities:
Infrastructure & Reliability: Design, deploy, and maintain scalable and highly available infrastructure. Own reliability, availability, performance, capacity, and incident response. Define and manage SLOs, SLIs, error budgets, and reliability practices.
CI/CD & Automation: Build and maintain CI/CD pipelines with automation, testing, and security gates. Improve deployment reliability, speed, and developer experience.
Containers & Kubernetes: Manage production containerized workloads and orchestration platforms using Docker and Kubernetes. Maintain secure, scalable, and highly available Kubernetes environments.
Infrastructure as Code: Design and maintain infrastructure using Terraform, Ansible, Helm, and other Infrastructure as Code and configuration-management tools.
Monitoring & Observability: Build and operate monitoring, logging, alerting, and observability platforms using technologies such as Prometheus, Grafana, ELK/OpenSearch, and Sentry. Monitor system health and optimize performance.
Incident Management & RCA: Troubleshoot complex production incidents, participate in on-call operations, lead root-cause analysis, and implement long-term reliability improvements.
Developer Collaboration: Work closely with developers to improve application operability, scalability, deployment processes, and overall engineering productivity.
Security & DevSecOps: Integrate security practices throughout infrastructure and CI/CD pipelines, including SAST, DAST, dependency and container image scanning, secrets management, and secure configuration.
Infrastructure Security: Harden Kubernetes, containers, Linux systems, and network infrastructure. Experience with WAF, IDS/IPS, DDoS protection, network security controls, Vault, KMS, and Sealed Secrets is preferred.
Required Skills & Experience:
Experience: 3–7+ years of experience in DevOps, SRE, Platform Engineering, or similar roles.
Technical Expertise: Strong Linux and networking fundamentals, production Kubernetes and container experience, CI/CD platforms such as GitLab CI, Jenkins, or Argo CD, and Infrastructure as Code using Terraform, Ansible, or Helm.
Monitoring & Observability: Hands-on experience with Prometheus, Grafana, ELK/OpenSearch, Sentry, or equivalent monitoring and observability platforms.
Automation: Strong scripting and automation skills using Bash, Python, or Go.
Security: Understanding of DevSecOps, least-privilege access, secrets management, system hardening, threat modeling, and Zero Trust concepts.
Nice to Have:
Security certifications or hands-on security operations experience, experience operating production systems at scale, and experience with incident management, on-call rotations, and reliability engineering practices.
What We Offer:
Ownership of critical production systems, opportunity to shape DevOps/SRE and DevSecOps practices, and a competitive and flexible working environment.
اکو هلدینگ با هدف ارائه راهکارهای متنوع در حوزه سرمایهگذاری و تسهیل آن از سال 94 شروع به فعالیت کرده است و دارای چهار زیرمجموعه تخصصی می باشد. «اکواحسان» بهعنوان رسانه تحلیلی و آموزشی بازارهای مالی، نقش مشاور سرمایهگذاری آگاهانه را ایفا میکند و محتوای کاربردی برای تصمیمگیری هوشمند مالی تولید مینماید. «اکوگلد» بستری امن و شفاف برای خرید و فروش طلای آبشده بدون واسطه فراهم کرده تا سرمایهگذاران با اطمینان و سهولت بیشتر وارد این بازار شوند. همچنین «اکوتراست» با تمرکز بر مشاوره سرمایهگذاری شخصیسازیشده، به افراد کمک میکند تا استراتژی مالی متناسب با اهداف خود را طراحی و اجرا کنند.