Senior Site Reliability Engineer, Infrastructure Foundations

Wikimedia Foundation

Completely RemoteFull TimeInformation Technology
Posted Today

Job description

Responsibilities

  • Perform day-to-day operational and DevOps tasks on public-facing infrastructure
  • Implement and utilize configuration management and deployment tools like Puppet and Kubernetes
  • Automate the installation, configuration, and maintenance of platform services
  • Assist product teams with architectural design for scalable functionality
  • Participate in a 24/7 on-call rotation for incident response and diagnosis
  • Collaborate with a global, cross-functional team in an asynchronous environment
  • Mentor peers in technical and operational areas

Requirements

  • 6+ years of experience in an SRE, Operations, or DevOps role
  • Proficiency with shell scripting and languages such as Python, Go, Bash, or Ruby
  • Experience with configuration management tools like Puppet or Ansible
  • Experience designing and managing infrastructure security for large service fleets
  • Experience with technical response during security incidents
  • Strong Linux system-level troubleshooting skills and package management (Debian)
  • Proven history of automating tasks and identifying process gaps
  • Strong English language skills for working in a distributed global team

Preferred Qualifications

  • Experience setting fleet-wide security policies and software supply chain security
  • Knowledge of monitoring, metrics, and logging infrastructure (Prometheus, Grafana)
  • Experience with LAMP stack technologies (PHP/HHVM, memcached/Redis)
  • Experience defining and implementing cross-team SLOs
  • Contributions to Free and Open Source software projects

About the Company

The Wikimedia Foundation is the nonprofit organization that operates Wikipedia and other free knowledge projects. Our vision is a world in which every single human can freely share in the sum of all knowledge.

Skills & tools

PythonKubernetesPuppet

What the team is looking for

Use this list as a quick fit check before you apply.

  1. 016+ years SRE/DevOps experience
  2. 02Python, Go, Bash, or Ruby proficiency
  3. 03Puppet or Ansible experience
  4. 04Linux system troubleshooting
  5. 05Infrastructure security management
  6. 06Incident response experience
NeverApplyAd

Wake up to a shortlist, not a search results page.

NeverApply scores every new listing against your CV, salary floor and visa. A handful of real matches by morning.

Get your daily matches