Site Reliability Engineer Jobs

64 open positions found · Salary range: $22 - $300,000

View Site Reliability Engineer
O

Site Reliability Engineer

OneStream SoftwareNot specifiedremote

From $401,000/yr

Location: Remote, United States Employment Type: Full-Time Benefits Offered: Vision, Medical, Life, Dental, 401K Gross Annual Base Salary: USD 114,000 - 148,000 Additional variable compensation and benefits may apply. Total compensation is based on experience, skills, and location using objective, job-related criteria. Summary As a Site Reliability Engineer, you will focus on ensuring the platform and services customers rely on are reliable, performant, and highly available. If you enjoy staying at the forefront of technology and automating infrastructure deployments, then this is the job for you. This vital role within Cloud Services requires knowledge and experience designing, implementing, and monitoring scalable and secure cloud services. The employee is expected to work well in a small team and willing to share responsibilities with other team members as needed. You will interact with internal staff, managers, and customers to implement and maintain operations. A passion for technology and learning, and the ability to grow others are vital for success in this role. Primary Duties And Responsibilities Implement application/infrastructure observability solutions to ensure desired application availability, reliability, and performance. Participate in regular On-Call rotations and share details related to incidents and their resolution through post-mortem reports and regular review meetings. Proactively partner with Product and Engineering teams to identify, develop, deploy, and maintain reliable systems and services. Influence and create new designs, architectures, standards, and methods for large-scale systems. Sustain a high level of reliability for key services and automated systems. Automate processes to improve reliability, performance, and availability. Update technical documentation, workflows, and knowledge base articles. Provide feedback in pull requests and peer coding reviews. Implement codified automated solutions that build integrations between Dynatrace, Azure DevOps and Jira. Solid knowledge in focused areas of OneStream Software. Ability to mentor others in several technical areas. Understanding practical use of SOC/FedRAMP controls to assist Compliance and Security teams. Required Education And Experience BS/BA in computer science, engineering, or technology-related field (or equivalent work experience). Proven work experience as a Site Reliability Engineer or in a similar role. 6+ years of cloud infrastructure and software development experience. 2+ years hands on experience of Azure Kubernetes Services (AKS) with container-based deployment skills or other platforms such as OpenShift, GKS, EKS. Advanced understanding of APM and observability tools such as Dynatrace, AppInsights, DataDog, Log Analytics, New Relic, Prometheus and Grafana. Advanced understanding of Infrastructure-as-Code (IaC) concepts and tooling (Terraform, CloudFormation templates, Bicep or ARM templates) on Microsoft Azure, Amazon Web Services (AWS), or Google Cloud Platform (GCP). Deep knowledge of Configuration Management/Orchestration utilities such as Ansible, PowerShell DSC, Chef, and Puppet. Advanced understanding of cloud concepts including elasticity, security, and identity management. Well versed familiarity with Agile Development methodologies utilizing Jira or Azure DevOps Boards. 6+ years of hands-on experience with the following technologies, tools, and concepts: + Automating processes using PowerShell, Bash, CLI, REST APIs, python, ARM Templates or other scripting languages. + Comfortable leveraging source control tools such as Git, Azure DevOps, or GitHub. + Knowledge of container orchestration platforms such as Kubernetes, OpenShift, AKS, GKS or helm. + Microsoft Azure, Amazon Web Services (AWS) or Google Cloud (GCP). Preferred Education And Experience Experience working for a cloud service provider (CSP), managed service provider (MSP), or SaaS provider. 6+ years of relevant Azure experience deploying and managing leveraging Infrastructure-as-Code (IAC) concepts. Experience with Microsoft and .NET (.NET, C#, SQL). Experience writing efficient and reliable code in a development environment. Debian, Ubuntu, Alpine or other distributions of the Linux operating systems. Deep knowledge and understanding of containerized applications, with special attention to reliability and monitoring of those containerized applications. Knowledge, Skills, And Abilities Deal well with ambiguous/undefined problems. Ability to self-motivate and work independently. Strong organizational and prioritization skills. Ability to find and apply effective solutions to emerging problems and challenges. Strong attention to detail. Comfortable communicating with all levels of management and engineering. Ability to get up to speed quickly with modern technologies and services. Ability to multitask on a variety of projects. Travel Travel Requirement: Travel is not expected to exceed 5%. Who We Are OneStream is how today’s Finance teams can go beyond just reporting on the past and Take Finance Further™ by steering the business to the future. It’s the only enterprise finance platform that unifies financial and operational data, embeds AI for better decisions and productivity, and empowers the CFO to become a critical driver of business strategy and execution. Our vision is to be the operating system for modern finance, digitizing core financial functions and empowering the CFO to become a critical driver of business strategy. To learn more visit www.onestream.com. Why Join The OneStream Team Transparency around corporate structure, salary, and benefits. Core value of customer success. Variety of project work (not industry-specific). Strong culture and camaraderie. Multiple training opportunities. Benefits At OneStream OneStream employees are passionate, hardworking individuals who go above and beyond to keep our customers happy and follow through on our mission statement. They consistently deliver the best and in turn, we make every effort to keep them cared for and happy. A sample of the benefits we provide are: Excellent Medical Plan. Dental \& Vision Insurance. Life Insurance. Short \& Long Term Disability. Vacation Time. Paid Holidays. Professional Development. Retirement Plan. All candidates must be legally authorized to work for any company in the country where this position is located without sponsorship. OneStream is an Equal Opportunity Employer.

via universal intelligenceApply ›
View Site Reliability Engineer
O

Site Reliability Engineer

OneStream SoftwareNot specifiedremote

From $401,000/yr

Location: Remote, United States Employment Type: Full-Time Benefits Offered: Vision, Medical, Life, Dental, 401K Gross Annual Base Salary: USD 114,000 - 148,000 Additional variable compensation and benefits may apply. Total compensation is based on experience, skills, and location using objective, job-related criteria. Summary As a Site Reliability Engineer, you will focus on ensuring the platform and services customers rely on are reliable, performant, and highly available. If you enjoy staying at the forefront of technology and automating infrastructure deployments, then this is the job for you. This vital role within Cloud Services requires knowledge and experience designing, implementing, and monitoring scalable and secure cloud services. The employee is expected to work well in a small team and willing to share responsibilities with other team members as needed. You will interact with internal staff, managers, and customers to implement and maintain operations. A passion for technology and learning, and the ability to grow others are vital for success in this role. Primary Duties And Responsibilities Implement application/infrastructure observability solutions to ensure desired application availability, reliability, and performance. Participate in regular On-Call rotations and share details related to incidents and their resolution through post-mortem reports and regular review meetings. Proactively partner with Product and Engineering teams to identify, develop, deploy, and maintain reliable systems and services. Influence and create new designs, architectures, standards, and methods for large-scale systems. Sustain a high level of reliability for key services and automated systems. Automate processes to improve reliability, performance, and availability. Update technical documentation, workflows, and knowledge base articles. Provide feedback in pull requests and peer coding reviews. Implement codified automated solutions that build integrations between Dynatrace, Azure DevOps and Jira. Solid knowledge in focused areas of OneStream Software. Ability to mentor others in several technical areas. Understanding practical use of SOC/FedRAMP controls to assist Compliance and Security teams. Required Education And Experience BS/BA in computer science, engineering, or technology-related field (or equivalent work experience). Proven work experience as a Site Reliability Engineer or in a similar role. 6+ years of cloud infrastructure and software development experience. 2+ years hands on experience of Azure Kubernetes Services (AKS) with container-based deployment skills or other platforms such as OpenShift, GKS, EKS. Advanced understanding of APM and observability tools such as Dynatrace, AppInsights, DataDog, Log Analytics, New Relic, Prometheus and Grafana. Advanced understanding of Infrastructure-as-Code (IaC) concepts and tooling (Terraform, CloudFormation templates, Bicep or ARM templates) on Microsoft Azure, Amazon Web Services (AWS), or Google Cloud Platform (GCP). Deep knowledge of Configuration Management/Orchestration utilities such as Ansible, PowerShell DSC, Chef, and Puppet. Advanced understanding of cloud concepts including elasticity, security, and identity management. Well versed familiarity with Agile Development methodologies utilizing Jira or Azure DevOps Boards. 6+ years of hands-on experience with the following technologies, tools, and concepts: + Automating processes using PowerShell, Bash, CLI, REST APIs, python, ARM Templates or other scripting languages. + Comfortable leveraging source control tools such as Git, Azure DevOps, or GitHub. + Knowledge of container orchestration platforms such as Kubernetes, OpenShift, AKS, GKS or helm. + Microsoft Azure, Amazon Web Services (AWS) or Google Cloud (GCP). Preferred Education And Experience Experience working for a cloud service provider (CSP), managed service provider (MSP), or SaaS provider. 6+ years of relevant Azure experience deploying and managing leveraging Infrastructure-as-Code (IAC) concepts. Experience with Microsoft and .NET (.NET, C#, SQL). Experience writing efficient and reliable code in a development environment. Debian, Ubuntu, Alpine or other distributions of the Linux operating systems. Deep knowledge and understanding of containerized applications, with special attention to reliability and monitoring of those containerized applications. Knowledge, Skills, And Abilities Deal well with ambiguous/undefined problems. Ability to self-motivate and work independently. Strong organizational and prioritization skills. Ability to find and apply effective solutions to emerging problems and challenges. Strong attention to detail. Comfortable communicating with all levels of management and engineering. Ability to get up to speed quickly with modern technologies and services. Ability to multitask on a variety of projects. Travel Travel Requirement: Travel is not expected to exceed 5%. Who We Are OneStream is how today’s Finance teams can go beyond just reporting on the past and Take Finance Further™ by steering the business to the future. It’s the only enterprise finance platform that unifies financial and operational data, embeds AI for better decisions and productivity, and empowers the CFO to become a critical driver of business strategy and execution. Our vision is to be the operating system for modern finance, digitizing core financial functions and empowering the CFO to become a critical driver of business strategy. To learn more visit www.onestream.com. Why Join The OneStream Team Transparency around corporate structure, salary, and benefits. Core value of customer success. Variety of project work (not industry-specific). Strong culture and camaraderie. Multiple training opportunities. Benefits At OneStream OneStream employees are passionate, hardworking individuals who go above and beyond to keep our customers happy and follow through on our mission statement. They consistently deliver the best and in turn, we make every effort to keep them cared for and happy. A sample of the benefits we provide are: Excellent Medical Plan. Dental \& Vision Insurance. Life Insurance. Short \& Long Term Disability. Vacation Time. Paid Holidays. Professional Development. Retirement Plan. All candidates must be legally authorized to work for any company in the country where this position is located without sponsorship. OneStream is an Equal Opportunity Employer.

via universal intelligenceApply ›
View DevOps/Site Reliability Engineer
B

DevOps/Site Reliability Engineer

Please read the note before you apply!!!!! Job Title: DevOps Engineer / Site Reliability Engineer (SRE) (W2) Company: Baanyan Software Services Inc Experience: 7+ Years Employment Type: Full-Time, W2 (No C2C Resumes Please) Location: Open to Relocation (Multiple Client Locations Across the US) Visa: Graduates with a Master's Degree / Bachelor's Degree Sponsorship: Yes Job Title: DevOps Engineer / Site Reliability Engineer (SRE) Job Description Responsibilities: Design, implement, and maintain scalable, highly available, and secure cloud infrastructure. Automate infrastructure provisioning, deployment, monitoring, and incident response processes. Build and manage CI/CD pipelines for application deployments across multiple environments. Manage containerized workloads using Docker and Kubernetes in cloud-native environments. Collaborate with development, QA, and security teams to improve software delivery and operational excellence. Monitor system performance, troubleshoot production issues, and ensure high availability and reliability. Implement Infrastructure as Code (IaC) using Terraform, CloudFormation, or similar tools. Manage logging, monitoring, alerting, and observability platforms. Ensure security, compliance, and best practices across cloud and infrastructure environments. Participate in on-call rotations and incident management activities. Optimize cloud infrastructure costs and improve system performance. Required Skills: 7+ years of hands-on experience in DevOps, Site Reliability Engineering (SRE), or Cloud Engineering. Strong experience with AWS, Azure, or Google Cloud Platform (GCP). Expertise in Kubernetes and Docker containerization technologies. Hands-on experience with Infrastructure as Code (Terraform, CloudFormation, Ansible). Strong experience building and maintaining CI/CD pipelines using Jenkins, GitHub Actions, GitLab CI/CD, or Azure DevOps. Experience with Linux/Unix administration and shell scripting (Bash, Python). Strong understanding of networking concepts, load balancing, DNS, SSL/TLS, VPNs, and security best practices. Experience with monitoring and observability tools such as Prometheus, Grafana, Datadog, ELK Stack, Splunk, New Relic, or Dynatrace. Experience with source control systems such as Git and GitHub/GitLab. Strong troubleshooting, incident management, and root cause analysis skills. Experience supporting highly available, distributed production systems. Familiarity with Agile and DevOps methodologies. Preferred Skills: Experience with service mesh technologies such as Istio or Linkerd. Experience with Kubernetes Operators and Helm Charts. Knowledge of AWS EKS, ECS, Lambda, EC2, RDS, S3, CloudWatch, IAM, and VPC. Experience implementing DevSecOps practices and security automation. Familiarity with SRE principles, including SLI, SLO, SLA, error budgets, and reliability engineering. Experience with Kafka, RabbitMQ, or other messaging platforms. Hands-on experience with disaster recovery, backup strategies, and business continuity planning. Certifications such as AWS Certified DevOps Engineer, AWS Solutions Architect, CKA, CKAD, or Terraform Associate. Experience using AI-powered development and operations tools such as GitHub Copilot, Windsurf, Cursor AI, Amazon Q, ChatGPT, and AI-driven observability platforms. Technologies: Cloud: AWS, Azure, GCP Containers: Docker, Kubernetes, OpenShift CI/CD: Jenkins, GitHub Actions, GitLab CI/CD, Azure DevOps IaC: Terraform, CloudFormation, Ansible Monitoring: Prometheus, Grafana, Datadog, Splunk, ELK Stack, New Relic Scripting: Python, Bash, PowerShell Source Control: Git, GitHub, GitLab, Bitbucket Operating Systems: Linux, Unix, Windows Server Messaging: Kafka, RabbitMQ, ActiveMQ Security: IAM, Vault, DevSecOps, Security Scanning Tools Thanks \& Regards Deepthi Gowada Talent Acquisition—Team Member Baanyan Software Services Inc 100 Metroplex Drive, Suite 100, 1st Floor, Edison, NJ. 08817 Phone: 732-660-9078 Extn: 208 Email: deepthi.g@baanyan.com | www.baanyan.com An E-Verified Company Pay: $40.00 - $45.00 per year Benefits: * Relocation assistance Work Location: In person

3 months agovia universal intelligenceApply ›
View Senior Site Reliability Engineer (Auth0)
O

Senior Site Reliability Engineer (Auth0)

OktaBarcelona, Spainonsite

Get to know Okta Okta is The World’s Identity Company. We free everyone to safely use any technology, anywhere, on any device or app. Our flexible and neutral products, Okta Platform and Auth0 Platform, provide secure access, authentication, and automation, placing identity at the core of business security and growth. At Okta, we celebrate a variety of perspectives and experiences. We are not lookin

via universal intelligenceApply ›
View Senior Site Reliability Engineer, Compute
R

Senior Site Reliability Engineer, Compute

RobloxSan Mateo, CA, United Statesonsite

$195,780 - $242,100/yr

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people c

via universal intelligenceApply ›
View Staff Site Reliability Engineer-Federal, Security Clearance
Z

Staff Site Reliability Engineer-Federal, Security Clearance

ZscalerCrystal City, Virginia, USA

$119,000 - $170,000/yr

About Zscaler Zscaler is a pioneer and global leader in zero trust security. The world’s largest businesses, critical infrastructure organizations, and government agencies rely on Zscaler to secure users, branches, applications, data & devices, and to accelerate digital transformation initiatives. Distributed across more than 160 data centers globally, the Zscaler Zero Trust Exchange platform combined with advanced AI combats billions of

7 months agovia universal intelligenceApply ›
View Site Reliability Engineering Manager
O

Site Reliability Engineering Manager

OktaBarcelona, Spainonsite

Get to know Okta Okta is The World’s Identity Company. We free everyone to safely use any technology, anywhere, on any device or app. Our flexible and neutral products, Okta Platform and Auth0 Platform, provide secure access, authentication, and automation, placing identity at the core of business security and growth. At Okta, we celebrate a variety of perspectives and experiences. We are not lookin

via universal intelligenceApply ›
View Staff Site Reliability Engineer
O

Staff Site Reliability Engineer

OktaBengaluru, Indiaonsite

Get to know Okta Okta is The World’s Identity Company. We free everyone to safely use any technology, anywhere, on any device or app. Our flexible and neutral products, Okta Platform and Auth0 Platform, provide secure access, authentication, and automation, placing identity at the core of business security and growth. At Okta, we celebrate a variety of perspectives and experiences. We are not lookin

via universal intelligenceApply ›
View Staff Site Reliability Engineer-Federal, Security Clearance
Z

Staff Site Reliability Engineer-Federal, Security Clearance

ZscalerCrystal City, Virginia, USAonsite

$119,000 - $170,000/yr

About Zscaler Zscaler is a pioneer and global leader in zero trust security. The world’s largest businesses, critical infrastructure organizations, and government agencies rely on Zscaler to secure users, branches, applications, data & devices, and to accelerate digital transformation initiatives. Distributed across more than 160 data centers globally, the Zscaler Zero Trust Exchange platform combined with advanced AI combats billions of

via universal intelligenceApply ›
View Senior Site Reliability Engineer, Identity Platform
C

Senior Site Reliability Engineer, Identity Platform

CoinbaseRemote - USAremote

Ready to be pushed beyond what you think you’re capable of? At Coinbase, our mission is to increase economic freedom in the world. It’s a massive, ambitious opportunity that demands the best of us, every day, as we build the emerging onchain platform — and with it, the future global financial system. To achieve our mission, we’re seeking a very specific candidate. We want someone who is passionate about our mission and who believes in the power of cryp

7 months agovia universal intelligenceApply ›
View Site Reliability Engineer - Platform
C

Site Reliability Engineer - Platform

CoinbaseRemote - USAremote

Ready to be pushed beyond what you think you’re capable of? At Coinbase, our mission is to increase economic freedom in the world. It’s a massive, ambitious opportunity that demands the best of us, every day, as we build the emerging onchain platform — and with it, the future global financial system. To achieve our mission, we’re seeking a very specific candidate. We want someone who is passionate about our mission and who believes in the power of cryp

7 months agovia universal intelligenceApply ›
View Director- Site Reliability Engineer
O

Director- Site Reliability Engineer

OktaBengaluru, Indiaonsite

Get to know Okta Okta is The World’s Identity Company. We free everyone to safely use any technology, anywhere, on any device or app. Our flexible and neutral products, Okta Platform and Auth0 Platform, provide secure access, authentication, and automation, placing identity at the core of business security and growth. At Okta, we celebrate a variety of perspectives and experiences. We are not lookin

via universal intelligenceApply ›
View Staff Site Reliability Engineer - Federal
Z

Staff Site Reliability Engineer - Federal

ZscalerCrystal City, Virginia, USA; Remote - Virginia, USAremote

About Zscaler Zscaler is a pioneer and global leader in zero trust security. The world’s largest businesses, critical infrastructure organizations, and government agencies rely on Zscaler to secure users, branches, applications, data & devices, and to accelerate digital transformation initiatives. Distributed across more than 160 data centers globally, the Zscaler Zero Trust Exchange platform combined with advanced AI combats billions of

7 months agovia universal intelligenceApply ›
View Staff Site Reliability Engineer, Cloud Efficiency
O

Staff Site Reliability Engineer, Cloud Efficiency

OktaBengaluru, India

Get to know OktaOkta is The World’s Identity Company. We free everyone to safely use any technology, anywhere, on any device or app. Our flexible and neutral products, Okta Platform and Auth0 Platform, provide secure access, authentication, and automation, placing identity at the core of business security and growth.At Okta, we celebrate a variety of perspectives and experiences. We are not lookin

7 months agovia universal intelligenceApply ›
View Senior Manager of Site Reliability Engineering
G

Senior Manager of Site Reliability Engineering

GoDaddyArizona (hybrid role, not eligible in Alaska, Mississippi, North Dakota, or the Virgin Islands; also not considering candidates from California, Seattle, or NYC)remote

From $401,000/yr

Location Details: Arizona At GoDaddy the future of work looks different for each team. Some teams work in the office full-time, others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely. This is a hybrid position. You’ll divide your time between working remotely from your home and an office, so you should live within commuting distance. Hybrid teams may work in-office as much as a few times a week or as little as once a month or quarter, as decided by leadership. The hiring manager can share more about what hybrid work might look like for this team. This position is not eligible to be performed in Alaska, Mississippi, North Dakota, or the Virgin Islands. GoDaddy is not currently considering candidates for this role in California, Seattle, or NYC. Join our team… GoDaddy is seeking an exceptional Senior Manager of Site Reliability Engineering to lead our Public Key Infrastructure (PKI) operations. In this pivotal role, you

recentlyvia universal intelligenceApply ›
View Senior Manager, Site Reliability Engineering - Infrastructure Platform
O

Senior Manager, Site Reliability Engineering - Infrastructure Platform

OktaBellevue, Washington

Get to know OktaOkta is The World’s Identity Company. We free everyone to safely use any technology, anywhere, on any device or app. Our flexible and neutral products, Okta Platform and Auth0 Platform, provide secure access, authentication, and automation, placing identity at the core of business security and growth.At Okta, we celebrate a variety of perspectives and experiences. We are not lookin

7 months agovia universal intelligenceApply ›
View Staff Site Reliability Engineer-Federal, Security Clearance
Z

Staff Site Reliability Engineer-Federal, Security Clearance

ZscalerCrystal City, Virginia, USA

About Zscaler Zscaler is a pioneer and global leader in zero trust security. The world’s largest businesses, critical infrastructure organizations, and government agencies rely on Zscaler to secure users, branches, applications, data & devices, and to accelerate digital transformation initiatives. Distributed across more than 160 data centers globally, the Zscaler Zero Trust Exchange platform combined with advanced AI combats billions of

7 months agovia universal intelligenceApply ›
View Senior Site Reliability Engineer, Compute
R

Senior Site Reliability Engineer, Compute

RobloxSan Mateo, CA, United States

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people c

7 months agovia universal intelligenceApply ›
View Director- Site Reliability Engineer
O

Director- Site Reliability Engineer

OktaBengaluru, India

Get to know OktaOkta is The World’s Identity Company. We free everyone to safely use any technology, anywhere, on any device or app. Our flexible and neutral products, Okta Platform and Auth0 Platform, provide secure access, authentication, and automation, placing identity at the core of business security and growth.At Okta, we celebrate a variety of perspectives and experiences. We are not lookin

7 months agovia universal intelligenceApply ›
View Staff Site Reliability Engineer, Streaming
A

Staff Site Reliability Engineer, Streaming

AlpacaRemoteremote

From $24/yr

Staff Site Reliability Engineer, Streaming Who We Are: Alpaca is a US-headquartered self-clearing broker-dealer and brokerage infrastructure for stocks, ETFs, options, crypto, fixed income, 24/5 trading, and more. Our recent Series C funding round brought our total investment to over $170 million, fueling our ambitious vision. Amongst our subsidiaries, Alpaca is a licensed financial services company, serving hundreds of financial institutions across 40 countries with our institutional-grade APIs. This includes broker-dealers, investment advisors, wealth managers, hedge funds, and crypto exchanges, totalling over 6 million brokerage accounts. Our global team is a diverse group of experienced engineers, traders, and brokerage professionals who are working to achieve our mission of opening financial services to everyone on the planet. We're deeply committed to open-source contributions and fostering a vibrant community, continuously enhancing our award-winning, developer-friendly API and

Site Reliability EngineeringPerformance Engineering
recentlyvia universal intelligenceApply ›
View Algorithmic Trading, Senior Site Reliability Engineer (SRE)
B

Algorithmic Trading, Senior Site Reliability Engineer (SRE)

BTIGNew York, New York, United States

Create a Job Alert Level-up your career by having opportunities at BTIG sent directly to your inbox. Job Algorithmic Trading, Senior Site Reliability Engineer (SRE) New York, New York, United States Franchise Sales, Healthcare, Vice President/Director 2026 Full-Time Investment Banking Analyst San Francisco, California, United States Investment Banking, Senior Analyst/Associate, DCM Equity Research Associate, Medical Technology Chicago, New York Equity Research Senior Associate, Software New York, San Francisco ; San Francisco, California, United States Equity Research Supervisory Analyst Research, Equity Research Associate, Biotechnology Structured Products, Analyst Structured Products, Credit Sales Executive, Associate/Vice President Structured Products, Mortgage Origination Sales, Vice President Technology, Service Desk Associate Technology, Sr. PostgreSQL Database Engineer, Vice President Technology, Trading Floor Desktop Support Engineer

recentlyvia universal intelligenceApply ›
View Senior Site Reliability Engineer
R

Senior Site Reliability Engineer

RegrelloSeattle, San Francisco, New York or remote in US, Canada, Mexicoremote

From $40/yr

As a Senior Site Reliability Engineer at Regrello, you'll shape the developer platform, collaborate with customers, and ensure the reliability and security of infrastructure and applications. Company Regrello is a 40-person startup reimagining automation in supply chains, in which companies still communicate about $13T of annual shipments almost entirely via email. This is a $220-billion, Amazon-sized market opportunity that’s been largely overlooked for over 20 years. Our team has experience building billion-dollar companies and includes veterans from Oracle, Palantir, Apple, Amazon, and Microsoft. Our customers include some of the largest electronics manufacturers in the world, all of which are hungry for us to succeed. Regrello has been funded by Andreessen Horowitz, Tiger Global, Dell, Bloomberg, and Ram Sriram (first investor in Google). Product We are building a global operating network that finally enables supply-chain companies to collaborate within one platform. Our workflow e

AWSAzureCircleCIGCPGithub Actions+4 more
recentlyvia universal intelligenceApply ›
View Staff Site Reliability Engineer (SRE & Platform Reliability)
A

Staff Site Reliability Engineer (SRE & Platform Reliability)

AffirmRemote Polandremote

Affirm is reinventing credit to make it more honest and friendly, giving consumers the flexibility to buy now and pay later without any hidden fees or compounding interest.Affirm is reinventing credit to make it more honest and friendly, giving consumers the flexibility to buy now and pay later without any hidden fees or compounding interest. Site Reliability Engineering at Affirm is a small, yet crucial, team that helps ou

7 months agovia universal intelligenceApply ›
View Senior Machine Learning Site Reliability Engineer
P

Senior Machine Learning Site Reliability Engineer

PrimaMilan / Madrid / Londonremote

From $201/yr

Senior Machine Learning Site Reliability Engineer Milan / Madrid / London Engineering – Engineering / Permanent Employment / Remote Are you looking for a new challenge? Fancy helping us shape the future of motor insurance? Prima could be the place for you. Since 2015, we’ve been using our love of data and tech to rethink motor insurance and bring drivers a great experience at a great price. Our story began in Italy, where we’ve quickly become the number one online motor insurance provider. In fact, we’re trusted by over 5 million drivers. And now we’re expanding to help millions more drivers in the UK and Spain. To help fuel that growth, we need a Senior Machine Learning Site Reliability Engineer to join our Infrastructure team. This team is the beating heart of Prima. You’ll be joining over 300 engineers across software development, infrastructure, operations and security. Fueled by curiosity, experimentation and collaboration, you’ll help deliver scalable, impactful solutions that sh

SRE practices in productionAWS expertiseKubernetesnetworkingDNS+9 more
recentlyvia universal intelligenceApply ›
View Site Reliability Engineer
G

Site Reliability Engineer

GoDaddyIndia, remoteremote

From $401,000/yr

Location Details: India, remote At GoDaddy, the future of work looks different for each team. Some teams work in the office full-time, others have a hybrid arrangement (they work remotely some days and in the office some days) , and some work entirely remotely. This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings. Join Our Team GoDaddy is seeking a Site Reliability Engineer to join our team and play a crucial role in building and maintaining our cutting-edge internal tools. These tools enable thousands of developers to deploy seamlessly to production and efficiently manage their AWS infrastructure. As a Site Reliability Engineer (SRE), you will collaborate with multinational development teams to ensure the smooth deployment and ongoing maintenance of our platform. What you'll get to do... Your experience should include... We've got your back... We offer a range of total rewards

recentlyvia universal intelligenceApply ›
View Senior Site Reliability Engineer
A

Senior Site Reliability Engineer

AffirmRemote Spainremote

Affirm is reinventing credit to make it more honest and friendly, giving consumers the flexibility to buy now and pay later without any hidden fees or compounding interest.Affirm is looking for a Senior Software Engineer for our Cloud Compute team to play a pivotal role in ensuring the robust and scalable foundation of our entire platform. As a fully remote team based in Spain, we are the engine room that manages all of Affirm's Kubernetes clusters. Our

7 months agovia universal intelligenceApply ›
View Senior Site Reliability Engineer (Auth0)
O

Senior Site Reliability Engineer (Auth0)

OktaBarcelona, Spain

Get to know OktaOkta is The World’s Identity Company. We free everyone to safely use any technology, anywhere, on any device or app. Our flexible and neutral products, Okta Platform and Auth0 Platform, provide secure access, authentication, and automation, placing identity at the core of business security and growth.At Okta, we celebrate a variety of perspectives and experiences. We are not lookin

7 months agovia universal intelligenceApply ›
View Senior Site Reliability Engineer - Infrastructure
U

Senior Site Reliability Engineer - Infrastructure

UnderdogfantasyUnited States/Remoteremote

At Underdog, we make sports more fun. Our thesis is simple: build the best products and we’ll build the biggest company in the space, because there’s so much more to be built for sports fans. We’re just over five years in, and we’re one of the fastest-growing sports companies ever, most recently valued at $1.3B. And it’s still the early days. We’ve built and scaled multiple games and products across fantasy sports, sports betting, and prediction market

7 months agovia universal intelligenceApply ›
View Principal Site Reliability Engineer
O

Principal Site Reliability Engineer

OktaBengaluru, India

Get to know OktaOkta is The World’s Identity Company. We free everyone to safely use any technology, anywhere, on any device or app. Our flexible and neutral products, Okta Platform and Auth0 Platform, provide secure access, authentication, and automation, placing identity at the core of business security and growth.At Okta, we celebrate a variety of perspectives and experiences. We are not lookin

7 months agovia universal intelligenceApply ›
View Site Reliability Engineer Intern (Summer 2026)
O

Site Reliability Engineer Intern (Summer 2026)

OktaBellevue, Washington

Get to know OktaOkta is The World’s Identity Company. We free everyone to safely use any technology, anywhere, on any device or app. Our flexible and neutral products, Okta Platform and Auth0 Platform, provide secure access, authentication, and automation, placing identity at the core of business security and growth.At Okta, we celebrate a variety of perspectives and experiences. We are not lookin

7 months agovia universal intelligenceApply ›
View Senior Site Reliability Engineer (SRE & Platform Reliability
A

Senior Site Reliability Engineer (SRE & Platform Reliability

AffirmRemote Polandremote

Affirm is reinventing credit to make it more honest and friendly, giving consumers the flexibility to buy now and pay later without any hidden fees or compounding interest.Site Reliability Engineering at Affirm is a small, yet crucial, team that helps our Engineering partners to “Operate What They Own” with excellence to protect their customers’ experience. SRE accomplishes this through defining frameworks and best practices for operating applications,

7 months agovia universal intelligenceApply ›
View Senior Site Reliability Engineer (SRE & Platform Reliability)
A

Senior Site Reliability Engineer (SRE & Platform Reliability)

AffirmRemote Spainremote

Affirm is reinventing credit to make it more honest and friendly, giving consumers the flexibility to buy now and pay later without any hidden fees or compounding interest.Site Reliability Engineering at Affirm is a small, yet crucial, team that helps our Engineering partners to “Operate What They Own” with excellence to protect their customers’ experience. SRE accomplishes this through defining frameworks and best practices for operating applications,

7 months agovia universal intelligenceApply ›
View Staff Site Reliability Engineer
O

Staff Site Reliability Engineer

OktaBengaluru, India

Get to know OktaOkta is The World’s Identity Company. We free everyone to safely use any technology, anywhere, on any device or app. Our flexible and neutral products, Okta Platform and Auth0 Platform, provide secure access, authentication, and automation, placing identity at the core of business security and growth.At Okta, we celebrate a variety of perspectives and experiences. We are not lookin

7 months agovia universal intelligenceApply ›
View Intermediate Site Reliability Engineer, Tenant Scale: Tenant Services
G

Intermediate Site Reliability Engineer, Tenant Scale: Tenant Services

GitlabRemote, Americas; Remote, EMEAremote

GitLab is an open-core software company that develops the most comprehensive AI-powered DevSecOps Platform, used by more than 100,000 organizations. Our mission is to enable everyone to contribute to and co-create the software that powers our world. When everyone can contribute, consumers become contributors, significantly accelerating human progre

7 months agovia universal intelligenceApply ›
View Site Reliability Engineer
D

Site Reliability Engineer

DevRevNot specified

From $150/yr

Site Reliability Engineer DevRev At DevRev, we’re building the future of work with Computer – your AI teammate. Computer is not just another tool. It’s built on the belief that the future of work should be about genuine human connection and collaboration – not piling on more apps. Computer is the best kind of teammate: it amplifies your strengths, takes repetition and frustration out of your day, and gives you more time and energy to do your best work. How? Easy: it’s the only platform capable of… Complete data unification Most AI products focus on either structured data (like CRM records and support tickets), or unstructured data (like documents and emails). Computer AirSync connects everything, unifying all your data sources (like Google Workspace, Jira, Notion) into one AI-ready source of truth: Computer Memory. Powerful search, reasoning, and action Once connected to all your tools and apps, Computer is embedded in your full business context. It can find and summarize, sure. Even m

Infrastructure as Code (IaC...Kubernetes clusters managementCI/CD pipelines implementat...Cloud databases administrat...Monitoring, alerting, and o...+3 more
7 months agovia universal intelligenceApply ›
View Senior Site Reliability Engineer
N

Senior Site Reliability Engineer

NexthinkNot specifiedhybrid

From $200/yr

Nexthink is the leader in digital employee experience management software. The company provides IT leaders with unprecedented insight allowing them to see, diagnose and fix issues at scale impacting employees anywhere, with any application or network, before employees notice the issue. As the first solution to allow IT to progress from reactive problem solving to proactive optimization, Nexthink enables its more than 1,200 customers to provide better digital experiences to more than 15 million employees. Dual headquartered in Lausanne, Switzerland and Boston, Massachusetts, Nexthink has 9 offices worldwide. #LI-Hybrid Job Description At Nexthink, we empower our customers with industry-leading solutions to enable continuous improvement of employee experience. We deliver unmatched visibility across all environments, so IT teams can consistently see, diagnose, and fix digital workplace issues. As a SaaS provider, our commitment is to deliver a seamless, resilient, and scalable platform ar

AWSKubernetesTerraformPythonGo+9 more
recentlyvia universal intelligenceApply ›
View Senior Site Reliability Engineer, Backend (Reliability Engineering)
A

Senior Site Reliability Engineer, Backend (Reliability Engineering)

AffirmRemote Canadaremote

Affirm is reinventing credit to make it more honest and friendly, giving consumers the flexibility to buy now and pay later without any hidden fees or compounding interest.Site Reliability Engineering at Affirm is a small, yet crucial, team that helps our Engineering partners to “Operate What They Own” with excellence to protect their customers’ experience. SRE accomplishes this through defining frameworks and best practices for operating applications,

7 months agovia universal intelligenceApply ›
View Intermediate Site Reliability Engineer, Environment Automation
G

Intermediate Site Reliability Engineer, Environment Automation

GitlabRemote, EMEAremote

GitLab is an open-core software company that develops the most comprehensive AI-powered DevSecOps Platform, used by more than 100,000 organizations. Our mission is to enable everyone to contribute to and co-create the software that powers our world. When everyone can contribute, consumers become contributors, significantly accelerating human progre

7 months agovia universal intelligenceApply ›
View Intermediate Site Reliability Engineer, Database Operations
G

Intermediate Site Reliability Engineer, Database Operations

GitlabRemote, EMEAremote

From $100,000/yr

GitLab is an open-core software company that develops the most comprehensive AI-powered DevSecOps Platform, used by more than 100,000 organizations. Our mission is to enable everyone to contribute to and co-create the software that powers our world. When everyone can contribute, consumers become contributors, significantly accelerating human progre

7 months agovia universal intelligenceApply ›
View Lead Site Reliability Engineer (Infrastructure)

Lead Site Reliability Engineer (Infrastructure)

remoteremote

We are seeking a Lead Site Reliability Engineer (Infrastructure) to join our fast-moving VSaaS engineering organization. This role carries responsibility for technical leadership and operational execution of the Infrastructure SRE team. You will own the reliability, scalability, and operability of our shared platform and production systems, while shaping how reliability engineering and SRE practices are applied across the organization and mentoring senior and staff engineers. You will work closely with product engineering and platform teams to ensure a seamless developer experience, while setting standards, driving priorities, and leading by example during incidents and high-impact operational work. This role requires a strong technical background in cloud infrastructure, distributed systems, CI/CD, and GitOps, along with hands-on development experience in Golang and/or Python, to improve developer workflows, automation, and long-term system reliability. This is a remote role in the Un

cloud infrastructuredistributed systemsCI/CDGitOpsGolang+1 more
recentlyvia universal intelligenceApply ›
View Senior Software Engineer, Site Reliability
A

Senior Software Engineer, Site Reliability

AsanaWarsaw

Asana’s rapid growth brings new challenges in keeping our systems fast, reliable, and resilient. As our product evolves, we’re making a major investment in reliability – and building a brand new SRE team in Warsaw is a key part of that strategy. This is your chance to help shape it from day one. This isn’t a traditional “ops” role – we’re looking for strong software engineers who are passionate about building reliable, distributed systems. You’ll work closely with a small SRE team in S

7 months agovia universal intelligenceApply ›
View Senior Site Reliability Engineer
G

Senior Site Reliability Engineer

GoDaddyRemote, Romaniaremote

From $22/yr

Location Details: Remote, Romania. GoDaddy offers diverse work arrangements: full-time office, hybrid (remote and in-office), and fully remote options for each team. This role is remote, allowing you to work from the comfort of your home. There may be occasional visits to a GoDaddy office to attend team events or meetings. This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings. Join our team Do you feel ready to make a significant difference in the technology industry? Our dynamic Systems Engineering team is on the lookout for an exceptional Senior Site Reliability Engineer to become a part of our ambitious tech crew. If you have a passion for tackling complex issues, automating all possible tasks, and crafting flawless server setups, this role could be the perfect fit for you. You'll engage with state-of-the-art technology in an environment that values technical excellence and f

recentlyvia universal intelligenceApply ›
View Site Reliability Engineer II
A

Site Reliability Engineer II

AxonSeattle (hybrid schedule)hybrid

$115,500 - $184,800/yr

Site Reliability Engineer II Join Axon and be a Force for Good. At Axon, we’re on a mission to Protect Life. We’re explorers, pursuing society’s most critical safety and justice issues with our ecosystem of devices and cloud software. Like our products, we work better together. We connect with candor and care, seeking out diverse perspectives from our customers, communities and each other. Life at Axon is fast-paced, challenging and meaningful. Here, you’ll take ownership and drive real change. Constantly grow as you work hard for a mission that matters at a company where you matter. Your Impact As a contributor in the SRE organization, you are passionate about delivering solutions to the real-time problems our mission-critical cloud native services encounter. You are also obsessed about achieving the high quality and reliability our customers demand. You will work closely not only with the SRE organization, but your technical deliverables will reach the entire engineering organization

cloud platforms management ...managed languages (Python, ...Kubernetes platforms (AKS, ...CI/CD platformsobservability tools (APM, l...+1 more
recentlyvia universal intelligenceApply ›
View Senior Site Reliability Engineer II
I

Senior Site Reliability Engineer II

InstacartCanada - Remote (BC, ON, AB, or NS only)remote

We're transforming the grocery industry At Instacart, we invite the world to share love through food because we believe everyone should have access to the food they love and more time to enjoy it together. Where others see a simple need for grocery delivery, we see exciting complexity and endless opportunity to serve the varied needs of our community. We work to deliver an essential service that customers rely on to get their

7 months agovia universal intelligenceApply ›
View Senior Site Reliability Engineer - AI Platform
N

Senior Site Reliability Engineer - AI Platform

N26Berlin

About the opportunity We are seeking a Senior Site Reliability Engineer to join the Platform Engineering Domain in the AI Platform Team. The mission of Platform Engineering is to provide trusted, performant, self-service platforms that empower product teams to build "the bank the world loves to use." The AI Platform team contributes to this mission by creating scalable, secure, and compliant infrastructure solutions that support MLOps and GenAI capabilities.</

7 months agovia universal intelligenceApply ›
View Senior Site Reliability Engineer, Environment Automation
G

Senior Site Reliability Engineer, Environment Automation

GitlabRemote, Americas; Remote, Canadaremote

From $100,000/yr

GitLab is an open-core software company that develops the most comprehensive AI-powered DevSecOps Platform, used by more than 100,000 organizations. Our mission is to enable everyone to contribute to and co-create the software that powers our world. When everyone can contribute, consumers become contributors, significantly accelerating human progre

7 months agovia universal intelligenceApply ›

Showing 50 of 64 jobs. Sign up to see all results and set up alerts.

Companies hiring for Site Reliability Engineer

Site Reliability Engineer Salary Data

View detailed salary ranges from 20 listings with reported compensation ›

Track these jobs with Jobfu

Get AI-powered resume tailoring, application tracking, and job alerts for Site Reliability Engineer roles. Sign up free.

Sign up free