Infrastructure / Site Reliability Engineer (SRE)
200 $US/hFourchette indicative communiquée par Mercor
Publié le 7 octobre 2026 · Candidatures jusqu'au 7 novembre 2026
Je vous redirige vers la page officielle de Mercor. La candidature est gratuite et se fait en anglais.
Je touche une commission de la plateforme si un candidat que j'ai orienté est recruté. Cela ne change rien pour vous.
The Infrastructure or Site Reliability Engineer builds, operates, and scales enterprise-grade production systems. Candidates need strong professional experience with Kubernetes, AWS, and cloud-native infrastructure.
Description en anglais, telle que publiée par Mercor.
Mercor connects exceptional technical talent with leading organisations working on ambitious technology and AI initiatives. We are looking for experienced Infrastructure / Site Reliability Engineers (SREs) to join a full-time engagement focused on building and operating complex, enterprise-grade infrastructure. We are seeking engineers with strong hands-on experience building, operating, debugging, and scaling sophisticated production systems. The ideal candidate has worked extensively with Kubernetes, AWS, observability platforms such as Datadog, and modern infrastructure tooling.
This is a full-time opportunity, and candidates must be able to commit to full-time engagement.
What You'll Do
-
Build, operate, and improve highly available and scalable production infrastructure.
-
Manage and optimise Kubernetes-based production environments.
-
Design and maintain cloud infrastructure, primarily across AWS.
-
Improve system reliability, availability, scalability, and operational efficiency.
-
Build and maintain observability across infrastructure and applications using Datadog or similar platforms.
-
Investigate production incidents, perform root-cause analysis, and implement durable fixes.
-
Improve monitoring, alerting, logging, tracing, and overall production visibility.
-
Develop automation and internal tooling to reduce manual operational work.
-
Partner closely with software engineering teams on deployments, infrastructure, and production reliability.
-
Contribute to infrastructure architecture and technical decisions for complex distributed systems.
Ideal Background
-
Professional experience in Infrastructure Engineering, Site Reliability Engineering (SRE), Platform Engineering, DevOps, or Production Engineering.
-
Hands-on experience operating complex, enterprise-grade production systems.
-
Strong production experience with Kubernetes.
-
Strong experience with AWS and cloud-native infrastructure.
-
Experience with Datadog, Prometheus, Grafana, or comparable observability platforms.
-
Experience with Infrastructure as Code using Terraform, Pulumi, or equivalent technologies.
-
Strong understanding of distributed systems, networking, containers, Linux, and cloud architecture.
-
Experience building or maintaining CI/CD and production deployment infrastructure.
-
Strong debugging, troubleshooting, and incident-response capabilities.
-
Proficiency in at least one programming or scripting language, such as Python, Go, or Bash.
Strong Signals
-
Experience operating Kubernetes and cloud infrastructure at significant production scale.
-
Experience supporting high-traffic or mission-critical applications.
-
Experience building infrastructure or platform tooling used by large engineering organisations.
-
Ownership of production reliability, on-call operations, incident response, or capacity planning.
-
Experience working within sophisticated, large-scale distributed systems.
-
Demonstrated improvements to SLOs/SLIs, observability, deployment reliability, infrastructure performance, or operational efficiency.
Why Join
-
Solve challenging reliability, scalability, and performance problems across enterprise-grade production systems.
-
Work extensively with technologies such as Kubernetes, AWS, Datadog, Terraform/Pulumi, and modern cloud-native tooling.
-
Take meaningful ownership of production reliability, observability, infrastructure architecture, and operational improvements.
-
Competitive hourly compensation reflecting your experience and technical expertise.
-
Join a network of highly skilled engineers working on ambitious projects with leading technology and AI organisations.
10 postes
Ouvert à l'international
À propos de Mercor
Mercor est une place de marché américaine qui recrute des experts à distance pour des projets d'IA et de conseil, du droit à l'ingénierie. L'offre originale est consultable sur leur site.
Voir l'offre sur MercorAutres postes en Développement logiciel
Builders & Developers: Share Your AI Coding Sessions
This remote role involves building or improving a real project using provided AI coding assistants while recording the session. Candidates must actively build projects with AI tools and be comfortable with simple terminal setups.
Publié aujourd'hui
Ingénieur QA, Applications 2D et Créatives
Ce rôle de contractuel à distance consiste à tester des applications de design 2D, d'animation et d'illustration pour évaluer des flux de travail réels et améliorer les systèmes d'IA. Les candidats doivent posséder une expérience professionnelle en assurance qualité, ainsi qu'une familiarité pratique avec le test de logiciels créatifs et le signalement de bugs.
- Quality Assurance (QA)
- Software testing
- 2D Design tools
- +3
Publié hier90 $US à 175 $US/h
Développeur Mobile Senior
Ce contractuel à distance crée des tâches de codage mobile complexes et des environnements d'apprentissage par renforcement pour entraîner des modèles d'IA avancés. Les candidats ont besoin d'une profonde expertise en développement mobile en Swift, Kotlin ou React Native, ainsi que de solides compétences en débogage.
- Kotlin
- Swift
- Mobile
Publié hier50 $US à 100 $US/h
Database Systems Evaluation Expert
The Database Systems Evaluation Expert designs realistic tasks and grading criteria to test how AI agents handle database engineering. Applicants must have at least five years of hands-on database engine internals experience or a database systems PhD.
Publié il y a 2 j
