About the Team
The Snapp! Doctor Infrastructure team is responsible for building, operating, and continuously improving the infrastructure and platform capabilities that support our products and services.
We work across areas such as service operations, databases, deployment platforms, observability, messaging, security, and business continuity. Our team owns the operational lifecycle of these capabilities from reliability and performance to security, resilience, and recovery. We collaborate closely with engineering, security, and other platform teams to keep our systems reliable, scalable, and ready for growth.
About the Role
As a member of the Snapp! Doctor Infrastructure team, you will work across a broad range of infrastructure and platform challenges in a production environment. You will contribute to operating and improving critical systems, automating repetitive work, improving reliability and observability, and supporting business continuity.
This role offers the opportunity to work with different technologies and problem spaces while taking meaningful ownership of the systems you support. You will gradually expand your responsibilities as you become familiar with our environment, with support from the team along the way.
We value ownership, knowledge sharing, continuous learning, and building systems and processes that do not depend on individual team members. If you enjoy solving complex infrastructure problems, learning by doing, and improving how technology is operated at scale, we'd like to hear from you.
Duties:
- Operate and maintain production infrastructure and platform services.
- Manage containerized workloads and Kubernetes-based environments.
- Build, maintain, and improve CI/CD and deployment workflows.
- Automate operational tasks through scripting and infrastructure-as-code.
- Maintain configuration and operational standards across environments, including monitoring, logging, alerting, and observability.
- Participate in the team’s on-call rotation and troubleshoot infrastructure and service issues and participate in incident response.
- Support backup, disaster recovery, and business continuity processes.
Requirements:
- Strong experience with Linux systems, networking fundamentals, and troubleshooting production environments.
- Hands-on experience with managing Kubernetes workers and containerized workloads using Helm.
- Good understanding of CI/CD principles and Experience with GitOps practices and tools such as ArgoCD
- Experience with monitoring, logging, and alerting systems such as Prometheus, Grafana, Splunk, or similar solutions.
- Good scripting skills with Bash and/or Python.
- Strong analytical and problem-solving skills, with the ability to identify the appropriate infrastructure or service layer when troubleshooting issues.
گروه اسنپ با برندهای شناختهشدهای همچون اسنپ، اسنپفود، اسنپباکس، اسنپتریپ، اسنپاستور، اسنپساپلای، اسنپدکتر، اسنپکیچن، اسنپپی، اسنپمارکت و اسنپشاپ شناخته میشود.
دستاوردهای چشمگیر گروه اسنپ، آن را به یکی از موفقترین کسبوکارهای ایران تبدیل کرده است.
ما به سرعت در حال رشد هستیم، و این به معنای فرصتهای نامحدود برای شماست.
به ما بپیوندید و در سفری هیجانانگیز در قلب توسعه کسبوکار و عضوی از یک تیم بینالمللی باشید.