Description
About the Role
The Site Reliability Engineering (SRE) team architects, builds, and maintains the rock-solid infrastructure that applications rely on. We work closely with development teams to ensure scalability, reliability, and efficiency. This collaboration empowers us to deliver exceptional customer experiences while enabling developers to focus on building great features.
What You Will Do
- Deploy, automate, maintain, and manage various cloud-based and on-premises production systems.
- Understanding the high-level overview of our architecture, and possessing the ability to systematically document new and existing requirements to ensure a smooth project delivery without miscommunication.
- Work closely with the Information security and infrastructure team in ensuring that we are adopting security best practices.
- Ensuring the availability, performance, scalability, and security of productions systems.
- Troubleshoot and resolve system issues across platform and application domains.
- Suggest architectural improvements and recommend process optimizations.
- Evaluate new technologies to enhance the infrastructure stack.
- Ensuring system security policies are properly remediated.
- Drive and implement automated provisioning and scaling of servers, along with testing and compliance checks using automation tools.
- Handle operational tasks, including on-call duties, alerts, and incident management.
What We Are Looking For
- At least 2 years of engineering experience.
- Bachelor’s or Master’s degree in a relevant field (e.g., IT, Computer Science) or a proven track record in DevOps.
- A strong willingness to continuously upgrade skills and stay up-to-date with the latest DevOps trends.
- Experience with cloud-native tools (e.g., Kubernetes, Docker, Nginx, OpenTelemetry) is a plus.
- Experience managing cloud servers (AWS, GCP).
- A desire to transition into engineering management is a valued addition.
- Experience with on-premises physical servers, databases, and storage solutions (MySQL, PostgreSQL, Redis) is a plus, as well as familiarity with Infrastructure as Code (IaC) tools (Terraform, Pulumi).
Similar jobs
About The Role The Site Reliability Engineering (SRE) team architects, builds, and maintains the rock-solid infrastructure that applications rely on. At the Senior Level, you own reliability, performance, and cost outcom…
Reolink, a leader in intelligent visual technology for homes and businesses, was founded in 2009 by a group of engineers with a strong commitment to and passion for smarter security solutions.Our products are now trusted…
Job Title: Site Reliability Engineer (SRE)Key Skills: Kubernetes, AWS/Azure/GCP, Terraform, Python, Observability, CI/CDExperience: +6 YOE.Location: Costa Rica, Peru, Colombia, and Bolivia.Mode: Remote. We at Coforge are…
Est. 90,000 GBP
Who are we? Ensono is a global technology services provider dedicated to helping organizations navigate the complexity of digital transformation. Through Ensono Product, Consulting & Technology, our dedicated consult…
Job Title: DevOps EngineerKey Skills: DevOps, Cloud Operations, AWS, TerraformLocation: BrazilMode: RemoteWe at Coforge are hiring DevOps Engineer with the following skill set.Key Responsibilities: Participate in a bi-we…
Reolink, a leader in intelligent visual technology for homes and businesses, was founded in 2009 by a group of engineers with a strong commitment to and passion for smarter security solutions.Our products are now trusted…
Jump Trading Group is committed to world class research. We empower exceptional talents in Mathematics, Physics, and Computer Science to seek scientific boundaries, push through them, and apply cutting edge research to g…
At OneSpan, we specialize in digital identity and anti-fraud solutions that create exceptional and secure experiences.We are looking for a Site Reliability Engineer to join our growing platform team in Delhi NCR. You wil…
Est. 82,000 USD
SummaryWe are seeking an experienced Site Reliability Engineering (SRE) / DevOps Manager to lead our teams and drive reliability roadmaps. As a key player in our leading crypto tax and portfolio tracking platform, you wi…
Est. 120,000 USD
SummaryWe are seeking an experienced Site Reliability Engineering (SRE) / DevOps Manager to lead our teams and drive reliability roadmaps. As a key player in our leading crypto tax and portfolio tracking platform, you wi…
We are representing a leading force in the decentralized exchange (DEX), and seeking a high-caliber technical leader to architect the backbone of a global financial ecosystem. In this role, you will bridge the gap betwee…
Est. 95,000 USD
SummaryWe are seeking an experienced Site Reliability Engineering (SRE) / DevOps Manager to lead our teams and drive reliability roadmaps. As a key player in our leading crypto tax and portfolio tracking platform, you wi…
Sr. Director, Site Reliability Engineering Coupang operates one of the largest and most complex technology platforms in the world. We are seeking a Senior Director, Site Reliability Engineering (Head of SRE) to define an…
Sr. Director, Site Reliability Engineering Coupang operates one of the largest and most complex technology platforms in the world. We are seeking a Senior Director, Site Reliability Engineering (Head of SRE) to define an…
Est. 70,000 USD
SummaryWe are seeking an experienced Site Reliability Engineering (SRE) / DevOps Manager to lead our teams and drive reliability roadmaps. As a key player in our leading crypto tax and portfolio tracking platform, you wi…
Jump Trading Group is committed to world class research. We empower exceptional talents in Mathematics, Physics, and Computer Science to seek scientific boundaries, push through them, and apply cutting edge research to g…
Summary: We are seeking an experienced Site Reliability Engineering (SRE) / DevOps Manager to lead our teams and drive reliability roadmaps. As a key player in our leading crypto tax and portfolio tracking platform, you…
Est. 180,000 USD
Summary: We are seeking an experienced Site Reliability Engineering (SRE) / DevOps Manager to lead our teams and drive reliability roadmaps. As a key player in our leading crypto tax and portfolio tracking platform, you…
Est. 120,000 USD
We are representing a leading force in the decentralized exchange (DEX), and seeking a high-caliber technical leader to architect the backbone of a global financial ecosystem. In this role, you will bridge the gap betwee…
Est. 110,000 USD
Summary: We are seeking an experienced Site Reliability Engineering (SRE) / DevOps Manager to lead our teams and drive reliability roadmaps. As a key player in our leading crypto tax and portfolio tracking platform, you…
Est. 60,000 USD
SummaryWe are seeking an experienced Site Reliability Engineering (SRE) / DevOps Manager to lead our teams and drive reliability roadmaps. As a key player in our leading crypto tax and portfolio tracking platform, you wi…
Est. 132,000 USD
SummaryWe are seeking an experienced Site Reliability Engineering (SRE) / DevOps Manager to lead our teams and drive reliability roadmaps. As a key player in our leading crypto tax and portfolio tracking platform, you wi…
We are seeking a skilled and passionate Engineer to join our team to build and operate a Whole-of-Government (WoG) runtime platform. As a Site Reliability Engineer, you will be responsible for designing and operating Git…
Est. 124,000 USD
Application Support Engineer (Site Reliability Engineer) Location: USAJob Type: Full-Time, no visa sponsorship available Coforge is seeking a Senior Application Support Engineer (SRE) to join our dynamic team of consulta…
Est. 120,000 EUR
About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to…
Est. 140,000 USD
BeyondTrust is a place where you can bring your purpose to life through the work that you do, creating a safer world through our cybersecurity SaaS portfolio. Our culture of flexibility, trust, and continual learning mea…
Title: Senior Site Reliability Engineer - I, Product Area FocusLocation: Noida (Hybrid) Summary of role Own availability, the most important product feature, by continually striving for sustained operational excellence o…
Join New Era Technology, where People First is at the heart of everything we do. With a global team of over 3,000 professionals, we're committed to creating a workplace where everyone feels valued, empowered, and inspire…
Senior Site Reliability Engineer I Location San Jose, Costa Rica - Remote Summary of role Own availability, the most important product feature, by continually striving for sustained operational excellence of Sumo’s plane…
Title: Staff Site Reliability Engineer, Product Area FocusLocation: Noida / Bangalore (Hybrid) Summary of role Own availability, the most important product feature, by continually striving for sustained operational excel…