
Senior Site Reliability Engineer
Remote
RemoteFull timeMid level$53k – $120kPosted today
Apply with JobAssistAbout the role
Who you are
- Solid professional experience in SRE, DevOps, or Platform Engineering
- Solid hands-on Kubernetes: operating and scaling production clusters and container tooling (Docker) and its ecosystem
- Experience building and managing cloud infrastructure on AWS (or similar)
- Strong infrastructure-as-code practice with Terraform
- Experience with reliability frameworks: SLOs, SLIs, error budgets, alerting strategies
- Solid observability background: OpenTelemetry, Grafana/Prometheus or similar
- Proficiency with CI/CD (GitLab CI, GitHub Actions, or similar) and deployment automation
- Comfortable with Golang, Bash/scripting; broader programming a plus
- Practical, embedded use of AI in infra/ops/dev work, agentic workflows with concrete, observable results, not just familiarity with the tools
- Clear and thoughtful communication, especially in an async-first, global setting
- Proactive, curious, and comfortable taking ownership of challenges
- Collaborative and respectful across cultures, time zones, and backgrounds
- Experience with 1 back-end programming language (Elixir, Nodejs, Python, etc)
- Experience running and configuring Linux systems in a non-cloud environment
- Security knowledge and capabilities from a defensive and offensive standpoint
What the job involves
- As a Senior SRE at Remote, you'll work with a high degree of autonomy on complex reliability and platform problems, owning the plan and execution of features and projects within our SRE/Platform domain
- You'll contribute to the platform's architecture and reliability strategy, translating ambiguous requirements into robust, maintainable solutions and raise the technical bar of the engineers around you while collaborating closely with product and security teams in an async-first, fully remote environment
- You'll work AI-natively day to day and build reusable AI workflows that make the whole team faster and more reliable, not just yourself
- Lead solution discovery and delivery for reliability and infrastructure problems with real ambiguity, complexity, or scope. Autonomously, coordinating with other contributors where needed
- Contribute to the platform's architecture, tooling, and roadmap. Influence team priorities and advocate for technical initiatives
- Help define and operate reliability practices for our platform: SLOs/SLIs, error budgets, alerting, observability. Take responsibility for the team's operational stance, using support/incident metrics to shape technical strategy
- Resolve cross-team requests, identify systemic issues, and turn recurring ones into reusable fixes and runbooks rather than one-off answers
- Work AI-natively and operationalise it for the team: use agentic workflows by default; build reusable prompts, skills, and tooling embedded in the codebase so others ship faster, safely; design agent-ready systems (clean interfaces, good observability) that make AI-assisted changes easy to review. Establish shared standards and domain-level guardrails (secure-by-default patterns, CI protections, AI-assisted review practices)
- Mentor and give timely, actionable feedback to less-senior engineers; participate in hiring, onboarding, and RFC discussions
- Collaborate with Security on platform hardening and threat mitigation; contribute to capacity and cost-efficiency of the infrastructure
- Participate in incident response and on-call rotations to rapidly resolve issues and maintain system reliability
- You'll report to: SRE Team Lead
The application process
- Please fill out the form below and upload your CV with a PDF format
- We kindly ask you to submit your application and CV in English, as this is the standardised language we use here at Remote
- If you don’t have an up to date CV but you are still interested in talking to us, please feel free to add a copy of your LinkedIn profile instead
- Interview with recruiter
- Interview with HM
- (async) Infrastructure exercise (you're not expected to spend more than 2 - 4 hours)
- Interview with the team (without any manager in the call so you can really get to know the people and ask all you want to ask)
- Bar Raiser Interview
- Executive Interview
- Offer + Background check (Veremark & Remote)
Benefits
- Unlimited personal time off
- Paid parental leave
- Co-working allowance
- Flexible working hours
- Learning budget
- Mental health support
- Company stock options
- Home office setup
- Celebratory allowances
- Branded swag
- Other country-specific benefits
Millions of jobs, with real people getting hired every day
200,000+
New jobs added daily7,000,000+
Verified job listings1,000,000+
Tailored applications submittedFAQ
Questions, answered.
Click "Apply with JobAssist" – we tailor your resume and application to this role and submit it for your approval.
Yes. This role at Remote was screened before publishing – we confirmed the employer before listing it.
The employer didn't disclose a salary range for this listing. JobAssist shows pay whenever it's available.
This position can be done from anywhere, with no in-office requirement.
Yes – every application is tailored from your profile and this job's requirements, and you can review and edit before it's sent.