Jobgether
Team Leader, SRE

How your CV stacks up
Upload your CV to see how well it fits this job role
?%
Team Leader, SRE
This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Team Leader, SRE based in United Kingdom.
As Team Leader, SRE, you’ll lead a highly technical Site Reliability Engineering team responsible for the reliability of critical cloud infrastructure and developer platforms.
The role combines approximately 60% hands-on technical contribution with 40% people leadership, giving you meaningful influence over both engineering direction and team development.
You’ll work across Kubernetes, AWS, PostgreSQL, CI infrastructure, observability, security, and reliability practices.
You’ll help evolve a growing SRE function, strengthening SLOs, incident response, on-call practices, and operational excellence.
At the same time, you’ll own the career development, performance, hiring, and overall health of your team.
The environment is fully remote and asynchronous, requiring strong written communication, autonomy, and thoughtful prioritization.
This is an opportunity to shape reliability practices while remaining close enough to the technology to guide decisions and lead through complex incidents.
Accountabilities
- Lead and develop an SRE team, owning the full career lifecycle of direct reports, including onboarding, feedback, performance management, progression, and hiring.
- Coach engineers on both technical craft and interpersonal skills, creating clear opportunities for growth and development.
- Foster a healthy, collaborative team environment by understanding team dynamics, addressing conflicts constructively, and maintaining effective retrospective practices.
- Represent the SRE team across engineering and with senior leadership, communicating priorities, challenges, progress, and technical direction.
- Define and prioritize SRE objectives, balancing operational commitments with project delivery and protecting the team's focus.
- Own the support rotation and on-call model, ensuring sustainable operational coverage and effective incident response.
- Provide technical leadership across core infrastructure, including Kubernetes, AWS, PostgreSQL, DNS, TLS, CI infrastructure, and related platform services.
- Drive the evolution of reliability practices, including Service Level Objectives (SLOs), error budgets, incident management, observability, and post-incident improvements.
- Partner with Security teams on infrastructure threats, patching, security controls, audits, and compliance requirements.
- Oversee infrastructure-related vendor relationships, including renewals and commercial discussions in collaboration with senior leadership.
- Stay hands-on enough to review technical work, challenge architectural decisions, contribute to complex problems, and provide credible technical direction during incidents.
- Help mature the organization's reliability practices by identifying operational gaps and turning lessons from incidents into lasting engineering improvements.
- Balance short-term operational demands with longer-term platform investments and engineering goals.
Reasons to use Rodeo
I’m in my final year doing Economics and I don’t know whether to apply for grad schemes now or do a masters first. What do you think?
Honest answer — it depends on where you want to end up. A lot of top grad schemes (Big 4, civil service, banking) don’t need a masters. Let’s look at the ones you’d be competitive for now, and we can decide if a masters actually adds anything.
Also worth knowing: most autumn 2026 applications are open now. Timing matters more than you think.
Start with a chat, not a search bar
Grad scheme, placement, apprenticeship? Not sure what you want yet — that's fine. Your agent talks it through with you and turns "I have no idea" into a shortlist.
Graduate Consultant — 2026 Scheme
Why you're a good match
StrongYour economics background and your summer at a regional bank line up with what PwC looks for on the consulting scheme. Applications close in four weeks.
See breakdownIt searches the market for you
Every day your agent scans the market matching roles against what actually matters to you, not just keywords on a CV.
Why you're a good match
You’ve got the grades and the economics background, and your bank internship is exactly the experience this scheme looks for. Apply soon — deadlines close within the month.
Experience fit
Your summer at the bank plus your econometrics coursework map directly to the day-one responsibilities on this scheme — client modelling, market briefings, and deal support.
Only hits
No noise. No "maybe this fits." Just roles with a clear explanation of why they're right — and where to focus when applying.
Requirements
- Proven experience leading an SRE, infrastructure, platform, or DevOps engineering team, with direct responsibility for team members' growth, performance, and career progression.
- Strong people-management and coaching skills, with demonstrated ability to develop engineers and provide clear, constructive feedback.
- Experience hiring engineers and assessing both technical capability and broader engineering judgment.
- Strong conflict-resolution and team-dynamics skills, with the ability to build commitment around shared organizational goals.
- Deep hands-on experience in Site Reliability Engineering, DevOps, cloud infrastructure, or a closely related discipline.
- Production experience with Kubernetes, including operational troubleshooting, reliability, scaling, and real-world failure scenarios.
- Significant experience with AWS and cloud infrastructure at meaningful scale.
- Hands-on experience building, enabling, or scaling AI infrastructure.
- Strong understanding of observability principles and practices.
- Experience with Infrastructure as Code, particularly Terraform.
- Experience with CI/CD platforms such as GitLab CI, GitHub Actions, Jenkins, or comparable technologies.
- Strong knowledge of Docker and shell scripting.
- Experience owning or operating reliability practices including incident response, on-call, SLOs, error budgets, and post-incident improvement processes.
- Previous experience working in regulated environments and understanding the associated operational and compliance requirements.
- Exceptional prioritization skills, particularly when operational workload competes with project delivery.
- Strong written communication skills and comfort working in an asynchronous, globally distributed environment.
- Ability to build strong relationships across engineering and become a trusted partner for teams bringing reliability challenges forward.
- Experience with a backend programming language such as Elixir, Java, Clojure, Node.js, Python, or similar is a plus.
- Familiarity with modern observability technologies such as OpenTelemetry, distributed tracing, or Honeycomb is advantageous.
- Experience with PostgreSQL or Aurora operations, including performance, connection pools, and query optimization, is beneficial.
- Experience administering Linux systems outside cloud environments is a plus.
- Knowledge of infrastructure security from both defensive and offensive perspectives is advantageous.
- Familiarity with cloud cost management and FinOps is beneficial.
- Experience growing an engineering team from a small starting point, including establishing hiring standards, is a plus.
- Fluent English communication skills are required.
- Ability to work effectively in a fully remote and asynchronous environment.


Get help with your application
Your very own career expert that helps elevate your application to the next level.
Benefits
- Fully remote working environment.
- Flexible, asynchronous working model that allows you to organize your schedule around your life.
- Opportunity to work with a globally distributed engineering organization.
- Significant ownership over both technical direction and people development.
- Approximately 60% individual-contributor technical work and 40% leadership responsibilities.
- Opportunity to shape and mature an evolving SRE and reliability practice.
- Exposure to large-scale cloud infrastructure, Kubernetes, AWS, PostgreSQL, observability, CI/CD, security, and AI infrastructure.
- Flexible paid time off.
- Flexible working hours.
- 16 weeks of paid parental leave.
- Budget for coworking spaces, learning, wellness, and gym memberships.
- Mental health support services.
- Stock options.
- Home office budget and IT equipment.
- Annual salary range of $75,450–$169,700 USD, with actual compensation determined by factors such as location, experience, relevant skills, training, business needs, and market conditions.
- Compensation and benefits are structured according to location and local market considerations.
- Start date: As soon as possible.
- Location: Romania, with a fully remote working arrangement.
How Jobgether Works
We use an AI-powered matching process to ensure your application is reviewed quickly, objectively, and fairly against the role's core requirements. Our system identifies the top-fitting candidates, and this shortlist is then shared directly with the hiring company. The final decision and next steps (interviews, assessments) are managed by their internal team.
We appreciate your interest and wish you the best!
Why Apply Through Jobgether?
Data Privacy Notice: By submitting your application, you acknowledge that Jobgether will process your personal data to evaluate your candidacy and share relevant information with the hiring employer. This processing is based on legitimate interest and pre-contractual measures under applicable data protection laws (including GDPR). You may exercise your rights (access, rectification, erasure, objection) at any time.
We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.
“It took my CV and asked me questions relevant to understanding what kind of jobs to suggest for me. Suggestions were almost perfect. Jobs were exactly what I’ve been looking for.”
Jessica, London
Location