DevOps & Infrastructure
Remote.com
Remote.com is hiring an Engineering Manager to lead its Site Reliability Engineering team, part of Platform Engineering. Remote is a fully distributed company operating across six continents that helps organizations recruit, pay and manage international teams compliantly. The SRE team owns the core infrastructure that keeps the product reliable: Kubernetes, AWS, PostgreSQL, DNS, TLS, CI infrastructure and observability. This is a team-leader role split roughly 60 percent individual contribution and 40 percent leadership, with four direct reports. Reliability practices at Remote are maturing rather than mature: the SLO framework is live on initial teams and needs to be expanded organization-wide, and there is real work to do on balancing operational load against project delivery. The listing states the location as Anywhere globally, with current team coverage strongest in EMEA and APAC and candidates in the Americas especially welcome to fill the gap. Compensation is an annual salary range of USD 75,450 to 169,700 depending on location and experience, plus stock options, flexible paid time off, async-first flexible hours, 16 weeks paid parental leave, co-working, learning and wellness budgets, mental health support, and a home office budget. Start date is as soon as possible and applications are accepted on an ongoing basis (CV in English as PDF, or a LinkedIn profile).
4 more roles like this one.
Own the full career lifecycle for the SRE team including onboarding, feedback, performance assessment, progression and hiring. Look after team health, dynamics and retrospective practice, and act as the team spokesperson to engineering and senior leadership. Define, commit to and prioritize SRE goals and explain the rationale, and manage the support rotation and on-call model. Own core infrastructure (Kubernetes, AWS, PostgreSQL, DNS, TLS, CI) and drive the implementation of reliability practices such as SLOs, error budgets, incident response and observability. Partner with security on threats, patching, controls, audit and compliance, and manage vendor relationships with director support. Stay hands-on enough to review the team's work, challenge designs and lead incidents credibly.
Experience leading an SRE, infrastructure or platform engineering team with ownership of reports' growth, performance and career progression, including coaching, addressing underperformance early and hiring. Hands-on background in site reliability, DevOps or cloud infrastructure engineering with production Kubernetes experience and AWS at meaningful scale. Hands-on AI building, enablement and infrastructure scaling. Solid observability practices, Infrastructure as Code with Terraform, CI/CD systems (GitLab CI, GitHub Actions or Jenkins), Docker and shell scripting. Experience running reliability practices (incident response, on-call, SLOs, error budgets) and working in regulated environments. Exceptional prioritization and clear written communication, since most leadership happens in writing in an async, distributed environment. Nice to have: Elixir or other backend languages, OpenTelemetry and Honeycomb, PostgreSQL or Aurora operations, Linux administration, security, FinOps, and experience growing small teams.
Salary not disclosed
See roles like this one
A free account opens every listing on EnRoute Jobs, saves the ones worth keeping, and scores each against your skills.
Create a free accountAlready have one? Log in