
Site Reliability Engineer (SRE)
Social Discovery Group
WorldwideFully remoteFull TimeMid LevelInformation Technology
No salary statedPosted 20 days ago2w agoApply link checked 54 minutes ago
Why we think you can apply
- Hiring model
- Open worldwide · the employer states no country restriction
- Work from
- Anywhere, the UAE included
Posted 20 days ago · apply link checked 54 minutes ago
About the role
Responsibilities
- Own and improve production infrastructure reliability and stability
- Prepare, execute, and support deployments and infrastructure changes
- Build and maintain Infrastructure-as-Code solutions using Ansible and Terraform
- Support and optimize Kubernetes-based and containerized environments
- Develop automation scripts and internal operational tooling
- Monitor system health, investigate incidents, and proactively improve observability
- Participate in CI/CD improvements together with Development, QA, DevOps, and SRE teams
- Work with monitoring and alerting systems to reduce downtime and improve system performance
- Maintain technical documentation, runbooks, and operational procedures
- Support DNS, WAF, CDN, and caching infrastructure where required
Requirements
- 3+ years of experience in SRE, DevOps, System Administration, or Build/Release Engineering
- Strong Linux administration and troubleshooting skills
- Hands-on experience with Kubernetes and containerization technologies (Docker/Podman)
- Experience with CI/CD pipelines, preferably GitLab CI
- Practical experience with Infrastructure-as-Code and configuration management tools (Ansible and/or Terraform)
- Experience with observability and monitoring tools such as Prometheus, Grafana, Zabbix, or VictoriaMetrics
- Good understanding of networking fundamentals, DNS, HTTP/HTTPS, load balancing, and troubleshooting
- Experience with Git and modern software delivery workflows
- Ability to work independently, take ownership, and proactively improve infrastructure
- Fluent Russian level for technical documentation and team communication
Preferred Qualifications
- AWS or GCP experience
- RabbitMQ / AMQP experience
- Cloudflare, Akamai, WAF, CDN experience
- Experience with tracing and advanced observability tooling
Benefits
- Remote opportunity to work full-time
- 28 calendar days of vacation per year
- 7 wellness days per year
- Bonuses for successful applicant referrals
- 50% payment for professional training and conferences
- Corporate discount for English lessons
- Health benefits compensation up to $1,000 gross per year
- Workplace organization reimbursement up to $1,000 gross every 3 years
- Internal gamified gratitude system
What we look for
- 3+ years SRE or DevOps experience
- Linux administration skills
- Kubernetes and containerization experience
- CI/CD pipeline experience
- Infrastructure-as-Code experience
- Observability and monitoring tools experience
- Networking fundamentals knowledge
- Git experience
- Fluent Russian
KubernetesAnsibleTerraformDockerGitLab CIPrometheus
SponsorWake up to a shortlist, not a search results page.
NeverApply scores every new listing against your CV, salary floor and visa. A handful of real matches by morning.
Get your daily matchesSimilar jobsbased on title, category and scope

Social Discovery Group
SponsorWake up to a shortlist, not a search results page.
NeverApply scores every new listing against your CV, salary floor and visa. A handful of real matches by morning.
Get your daily matchesSite Reliability Engineer (SRE)Social Discovery Group · 3 free applies a month