Reliability Engineer Jobs

94 open positions found · Salary range: $12 - $300,000

View Site Reliability Engineer
O

Site Reliability Engineer

OneStream SoftwareNot specifiedremote

From $401,000/yr

Location: Remote, United States Employment Type: Full-Time Benefits Offered: Vision, Medical, Life, Dental, 401K Gross Annual Base Salary: USD 114,000 - 148,000 Additional variable compensation and benefits may apply. Total compensation is based on experience, skills, and location using objective, job-related criteria. Summary As a Site Reliability Engineer, you will focus on ensuring the platform and services customers rely on are reliable, performant, and highly available. If you enjoy staying at the forefront of technology and automating infrastructure deployments, then this is the job for you. This vital role within Cloud Services requires knowledge and experience designing, implementing, and monitoring scalable and secure cloud services. The employee is expected to work well in a small team and willing to share responsibilities with other team members as needed. You will interact with internal staff, managers, and customers to implement and maintain operations. A passion for technology and learning, and the ability to grow others are vital for success in this role. Primary Duties And Responsibilities Implement application/infrastructure observability solutions to ensure desired application availability, reliability, and performance. Participate in regular On-Call rotations and share details related to incidents and their resolution through post-mortem reports and regular review meetings. Proactively partner with Product and Engineering teams to identify, develop, deploy, and maintain reliable systems and services. Influence and create new designs, architectures, standards, and methods for large-scale systems. Sustain a high level of reliability for key services and automated systems. Automate processes to improve reliability, performance, and availability. Update technical documentation, workflows, and knowledge base articles. Provide feedback in pull requests and peer coding reviews. Implement codified automated solutions that build integrations between Dynatrace, Azure DevOps and Jira. Solid knowledge in focused areas of OneStream Software. Ability to mentor others in several technical areas. Understanding practical use of SOC/FedRAMP controls to assist Compliance and Security teams. Required Education And Experience BS/BA in computer science, engineering, or technology-related field (or equivalent work experience). Proven work experience as a Site Reliability Engineer or in a similar role. 6+ years of cloud infrastructure and software development experience. 2+ years hands on experience of Azure Kubernetes Services (AKS) with container-based deployment skills or other platforms such as OpenShift, GKS, EKS. Advanced understanding of APM and observability tools such as Dynatrace, AppInsights, DataDog, Log Analytics, New Relic, Prometheus and Grafana. Advanced understanding of Infrastructure-as-Code (IaC) concepts and tooling (Terraform, CloudFormation templates, Bicep or ARM templates) on Microsoft Azure, Amazon Web Services (AWS), or Google Cloud Platform (GCP). Deep knowledge of Configuration Management/Orchestration utilities such as Ansible, PowerShell DSC, Chef, and Puppet. Advanced understanding of cloud concepts including elasticity, security, and identity management. Well versed familiarity with Agile Development methodologies utilizing Jira or Azure DevOps Boards. 6+ years of hands-on experience with the following technologies, tools, and concepts: + Automating processes using PowerShell, Bash, CLI, REST APIs, python, ARM Templates or other scripting languages. + Comfortable leveraging source control tools such as Git, Azure DevOps, or GitHub. + Knowledge of container orchestration platforms such as Kubernetes, OpenShift, AKS, GKS or helm. + Microsoft Azure, Amazon Web Services (AWS) or Google Cloud (GCP). Preferred Education And Experience Experience working for a cloud service provider (CSP), managed service provider (MSP), or SaaS provider. 6+ years of relevant Azure experience deploying and managing leveraging Infrastructure-as-Code (IAC) concepts. Experience with Microsoft and .NET (.NET, C#, SQL). Experience writing efficient and reliable code in a development environment. Debian, Ubuntu, Alpine or other distributions of the Linux operating systems. Deep knowledge and understanding of containerized applications, with special attention to reliability and monitoring of those containerized applications. Knowledge, Skills, And Abilities Deal well with ambiguous/undefined problems. Ability to self-motivate and work independently. Strong organizational and prioritization skills. Ability to find and apply effective solutions to emerging problems and challenges. Strong attention to detail. Comfortable communicating with all levels of management and engineering. Ability to get up to speed quickly with modern technologies and services. Ability to multitask on a variety of projects. Travel Travel Requirement: Travel is not expected to exceed 5%. Who We Are OneStream is how today’s Finance teams can go beyond just reporting on the past and Take Finance Further™ by steering the business to the future. It’s the only enterprise finance platform that unifies financial and operational data, embeds AI for better decisions and productivity, and empowers the CFO to become a critical driver of business strategy and execution. Our vision is to be the operating system for modern finance, digitizing core financial functions and empowering the CFO to become a critical driver of business strategy. To learn more visit www.onestream.com. Why Join The OneStream Team Transparency around corporate structure, salary, and benefits. Core value of customer success. Variety of project work (not industry-specific). Strong culture and camaraderie. Multiple training opportunities. Benefits At OneStream OneStream employees are passionate, hardworking individuals who go above and beyond to keep our customers happy and follow through on our mission statement. They consistently deliver the best and in turn, we make every effort to keep them cared for and happy. A sample of the benefits we provide are: Excellent Medical Plan. Dental \& Vision Insurance. Life Insurance. Short \& Long Term Disability. Vacation Time. Paid Holidays. Professional Development. Retirement Plan. All candidates must be legally authorized to work for any company in the country where this position is located without sponsorship. OneStream is an Equal Opportunity Employer.

via universal intelligenceApply ›
View Site Reliability Engineer
O

Site Reliability Engineer

OneStream SoftwareNot specifiedremote

From $401,000/yr

Location: Remote, United States Employment Type: Full-Time Benefits Offered: Vision, Medical, Life, Dental, 401K Gross Annual Base Salary: USD 114,000 - 148,000 Additional variable compensation and benefits may apply. Total compensation is based on experience, skills, and location using objective, job-related criteria. Summary As a Site Reliability Engineer, you will focus on ensuring the platform and services customers rely on are reliable, performant, and highly available. If you enjoy staying at the forefront of technology and automating infrastructure deployments, then this is the job for you. This vital role within Cloud Services requires knowledge and experience designing, implementing, and monitoring scalable and secure cloud services. The employee is expected to work well in a small team and willing to share responsibilities with other team members as needed. You will interact with internal staff, managers, and customers to implement and maintain operations. A passion for technology and learning, and the ability to grow others are vital for success in this role. Primary Duties And Responsibilities Implement application/infrastructure observability solutions to ensure desired application availability, reliability, and performance. Participate in regular On-Call rotations and share details related to incidents and their resolution through post-mortem reports and regular review meetings. Proactively partner with Product and Engineering teams to identify, develop, deploy, and maintain reliable systems and services. Influence and create new designs, architectures, standards, and methods for large-scale systems. Sustain a high level of reliability for key services and automated systems. Automate processes to improve reliability, performance, and availability. Update technical documentation, workflows, and knowledge base articles. Provide feedback in pull requests and peer coding reviews. Implement codified automated solutions that build integrations between Dynatrace, Azure DevOps and Jira. Solid knowledge in focused areas of OneStream Software. Ability to mentor others in several technical areas. Understanding practical use of SOC/FedRAMP controls to assist Compliance and Security teams. Required Education And Experience BS/BA in computer science, engineering, or technology-related field (or equivalent work experience). Proven work experience as a Site Reliability Engineer or in a similar role. 6+ years of cloud infrastructure and software development experience. 2+ years hands on experience of Azure Kubernetes Services (AKS) with container-based deployment skills or other platforms such as OpenShift, GKS, EKS. Advanced understanding of APM and observability tools such as Dynatrace, AppInsights, DataDog, Log Analytics, New Relic, Prometheus and Grafana. Advanced understanding of Infrastructure-as-Code (IaC) concepts and tooling (Terraform, CloudFormation templates, Bicep or ARM templates) on Microsoft Azure, Amazon Web Services (AWS), or Google Cloud Platform (GCP). Deep knowledge of Configuration Management/Orchestration utilities such as Ansible, PowerShell DSC, Chef, and Puppet. Advanced understanding of cloud concepts including elasticity, security, and identity management. Well versed familiarity with Agile Development methodologies utilizing Jira or Azure DevOps Boards. 6+ years of hands-on experience with the following technologies, tools, and concepts: + Automating processes using PowerShell, Bash, CLI, REST APIs, python, ARM Templates or other scripting languages. + Comfortable leveraging source control tools such as Git, Azure DevOps, or GitHub. + Knowledge of container orchestration platforms such as Kubernetes, OpenShift, AKS, GKS or helm. + Microsoft Azure, Amazon Web Services (AWS) or Google Cloud (GCP). Preferred Education And Experience Experience working for a cloud service provider (CSP), managed service provider (MSP), or SaaS provider. 6+ years of relevant Azure experience deploying and managing leveraging Infrastructure-as-Code (IAC) concepts. Experience with Microsoft and .NET (.NET, C#, SQL). Experience writing efficient and reliable code in a development environment. Debian, Ubuntu, Alpine or other distributions of the Linux operating systems. Deep knowledge and understanding of containerized applications, with special attention to reliability and monitoring of those containerized applications. Knowledge, Skills, And Abilities Deal well with ambiguous/undefined problems. Ability to self-motivate and work independently. Strong organizational and prioritization skills. Ability to find and apply effective solutions to emerging problems and challenges. Strong attention to detail. Comfortable communicating with all levels of management and engineering. Ability to get up to speed quickly with modern technologies and services. Ability to multitask on a variety of projects. Travel Travel Requirement: Travel is not expected to exceed 5%. Who We Are OneStream is how today’s Finance teams can go beyond just reporting on the past and Take Finance Further™ by steering the business to the future. It’s the only enterprise finance platform that unifies financial and operational data, embeds AI for better decisions and productivity, and empowers the CFO to become a critical driver of business strategy and execution. Our vision is to be the operating system for modern finance, digitizing core financial functions and empowering the CFO to become a critical driver of business strategy. To learn more visit www.onestream.com. Why Join The OneStream Team Transparency around corporate structure, salary, and benefits. Core value of customer success. Variety of project work (not industry-specific). Strong culture and camaraderie. Multiple training opportunities. Benefits At OneStream OneStream employees are passionate, hardworking individuals who go above and beyond to keep our customers happy and follow through on our mission statement. They consistently deliver the best and in turn, we make every effort to keep them cared for and happy. A sample of the benefits we provide are: Excellent Medical Plan. Dental \& Vision Insurance. Life Insurance. Short \& Long Term Disability. Vacation Time. Paid Holidays. Professional Development. Retirement Plan. All candidates must be legally authorized to work for any company in the country where this position is located without sponsorship. OneStream is an Equal Opportunity Employer.

via universal intelligenceApply ›
View DevOps/Site Reliability Engineer
B

DevOps/Site Reliability Engineer

Please read the note before you apply!!!!! Job Title: DevOps Engineer / Site Reliability Engineer (SRE) (W2) Company: Baanyan Software Services Inc Experience: 7+ Years Employment Type: Full-Time, W2 (No C2C Resumes Please) Location: Open to Relocation (Multiple Client Locations Across the US) Visa: Graduates with a Master's Degree / Bachelor's Degree Sponsorship: Yes Job Title: DevOps Engineer / Site Reliability Engineer (SRE) Job Description Responsibilities: Design, implement, and maintain scalable, highly available, and secure cloud infrastructure. Automate infrastructure provisioning, deployment, monitoring, and incident response processes. Build and manage CI/CD pipelines for application deployments across multiple environments. Manage containerized workloads using Docker and Kubernetes in cloud-native environments. Collaborate with development, QA, and security teams to improve software delivery and operational excellence. Monitor system performance, troubleshoot production issues, and ensure high availability and reliability. Implement Infrastructure as Code (IaC) using Terraform, CloudFormation, or similar tools. Manage logging, monitoring, alerting, and observability platforms. Ensure security, compliance, and best practices across cloud and infrastructure environments. Participate in on-call rotations and incident management activities. Optimize cloud infrastructure costs and improve system performance. Required Skills: 7+ years of hands-on experience in DevOps, Site Reliability Engineering (SRE), or Cloud Engineering. Strong experience with AWS, Azure, or Google Cloud Platform (GCP). Expertise in Kubernetes and Docker containerization technologies. Hands-on experience with Infrastructure as Code (Terraform, CloudFormation, Ansible). Strong experience building and maintaining CI/CD pipelines using Jenkins, GitHub Actions, GitLab CI/CD, or Azure DevOps. Experience with Linux/Unix administration and shell scripting (Bash, Python). Strong understanding of networking concepts, load balancing, DNS, SSL/TLS, VPNs, and security best practices. Experience with monitoring and observability tools such as Prometheus, Grafana, Datadog, ELK Stack, Splunk, New Relic, or Dynatrace. Experience with source control systems such as Git and GitHub/GitLab. Strong troubleshooting, incident management, and root cause analysis skills. Experience supporting highly available, distributed production systems. Familiarity with Agile and DevOps methodologies. Preferred Skills: Experience with service mesh technologies such as Istio or Linkerd. Experience with Kubernetes Operators and Helm Charts. Knowledge of AWS EKS, ECS, Lambda, EC2, RDS, S3, CloudWatch, IAM, and VPC. Experience implementing DevSecOps practices and security automation. Familiarity with SRE principles, including SLI, SLO, SLA, error budgets, and reliability engineering. Experience with Kafka, RabbitMQ, or other messaging platforms. Hands-on experience with disaster recovery, backup strategies, and business continuity planning. Certifications such as AWS Certified DevOps Engineer, AWS Solutions Architect, CKA, CKAD, or Terraform Associate. Experience using AI-powered development and operations tools such as GitHub Copilot, Windsurf, Cursor AI, Amazon Q, ChatGPT, and AI-driven observability platforms. Technologies: Cloud: AWS, Azure, GCP Containers: Docker, Kubernetes, OpenShift CI/CD: Jenkins, GitHub Actions, GitLab CI/CD, Azure DevOps IaC: Terraform, CloudFormation, Ansible Monitoring: Prometheus, Grafana, Datadog, Splunk, ELK Stack, New Relic Scripting: Python, Bash, PowerShell Source Control: Git, GitHub, GitLab, Bitbucket Operating Systems: Linux, Unix, Windows Server Messaging: Kafka, RabbitMQ, ActiveMQ Security: IAM, Vault, DevSecOps, Security Scanning Tools Thanks \& Regards Deepthi Gowada Talent Acquisition—Team Member Baanyan Software Services Inc 100 Metroplex Drive, Suite 100, 1st Floor, Edison, NJ. 08817 Phone: 732-660-9078 Extn: 208 Email: deepthi.g@baanyan.com | www.baanyan.com An E-Verified Company Pay: $40.00 - $45.00 per year Benefits: * Relocation assistance Work Location: In person

3 months agovia universal intelligenceApply ›
View Lead Data Engineer – Modernization & Reliability
H

Lead Data Engineer – Modernization & Reliability

HumanaNew York, United Statesonsite

From $142,300/yr

Become a part of our caring community The Lead Data Engineer is responsible for leading the modernization, optimization, and stabilization of the Wisconsin Medicaid Market's data platform ecosystem across two independent health plan technology stacks. This role owns the market's Data Warehouse and ODS, drives ETL/data movement strategy (including SSIS modernization), and improves the reliability, observability, and security posture of data pipelines supporting critical Medicaid operations. The Lead Data Engineer partners closely with the Market BI team as their IT counterpart to improve data access and flow, including establishing and managing Databricks pipeline patterns and platform enablement as the BI environment evolves. The Lead Data Engineer owns and evolves the Wisconsin Medicaid Market's data stores and data movement ecosystem, including the Data Warehouse and ODS, and the ETL processes that connect vendor and internal systems. This role is accountable for modernizing and optimizing the market's data platform to improve reliability, reduce technical debt, strengthen observability/fault tolerance, and increase engineering efficiency in a complex dual-stack environment. The Lead Data Engineer is also the Market BI team's primary IT partner for data platform enablement — improving data access patterns, strengthening pipeline governance, and helping establish a maintainable approach to Databricks pipelines and workflows as the market matures its analytics capabilities. This is a Lead-level individual contributor role where work requires higher autonomy and complexity and includes directing the work of a small number of contract resources while remaining hands-on in delivery. Team Culture \& Expectations On the Wisconsin Medicaid Market IT team, how you show up is as important as what you accomplish. We value a positive attitude, curiosity, ownership, and a desire to learn and make a meaningful difference. This role requires collaboration, high accountability, and comfort operating in ambiguity while continuously improving data reliability, security, and outcomes for associates, members, providers, and other stakeholders. Skills \& Capabilities Use your skills to make an impact The ideal candidate blends deep data engineering fundamentals with the ability to lead through influence — setting direction, improving reliability, and guiding others' work (including contract resources) while staying hands-on. Required Skills * SQL, ETL/ELT \& Data Platform Depth - Proven production experience with SQL Server data engineering, including warehousing patterns, operational support, performance tuning, and ETL/ELT design. Strong experience with SSIS/SSRS and Azure Data Factory (or similar orchestration) in real operating environments. * Integration \& Interoperability Tooling (WI Market Reality) - Experience working across integration engines and healthcare data movement patterns, including tools such as Rhapsody / CorePoint * Modernization, Optimization \& Stabilization Mindset - Demonstrated ability to modernize and optimize fragile pipelines and legacy patterns, reduce technical debt, and improve reliability, observability, and fault tolerance. * Databricks / UDAP Growth Path (directional, not gatekeeping) - Experience with Databricks and/or enterprise data platforms (e.g., UDAP), or strong aptitude and desire to grow into these platforms as part of market modernization and enterprise alignment. * Execution Leadership - Highly organized, self-directed, and able to drive work to outcomes in an ambiguous, rapidly changing environment including planning, sequencing, and communicating progress/risk. Experience creating detailed technical documentation to gain buy-in and drive decisions. Team operates in SAFe Agile practices. * Vendor / Contract Resource Direction - Ability to direct and review contract resource work (clarify requirements, establish standards, review deliverables, ensure maintainability). Preferred Skills * Own the market's data stores (Data Warehouse + ODS) and ensure stability, security, and performance * AI Integration * Lead modernization and optimization of the ETL ecosystem, including SSIS modernization and improved reliability/observability * Drive reduction of legacy report patterns and support transition aligned to enterprise SSRS footprint reduction efforts * Partner with Market BI as their IT counterpart to improve data access and establish/manage Databricks pipeline patterns * Define and enforce data engineering standards aligned with enterprise direction. * Reduce technical debt and streamline workflows to lower operational burden on engineers and increase delivery capacity Additional Information Work-At-Home Requirements * WAH requirements: Must have the ability to provide a high speed DSL or cable modem for a home office. Associates or contractors who live and work from home in the state of California will be provided payment for their internet expense. * A minimum standard speed for optimal performance of 25x10 (25mpbs download x 10mpbs upload) is required. * Satellite and Wireless Internet service is NOT allowed for this role. * A dedicated space lacking ongoing interruptions to protect member PHI / HIPAA information Travel: While this is a remote position, occasional travel to Humana's offices for training or meetings may be required. Scheduled Weekly Hours 40 Pay Range The compensation range below reflects a good faith estimate of starting base pay for full time (40 hours per week) employment at the time of posting. The pay range may be higher or lower based on geographic location and individual pay will vary based on demonstrated job related skills, knowledge, experience, education, certifications, etc. $142,300 - $195,700 per year This job is eligible for a bonus incentive plan. This incentive opportunity is based upon company and/or individual performance. Description Of Benefits Humana, Inc. and its affiliated subsidiaries (collectively, “Humana”) offers competitive benefits that support whole-person well-being. Associate benefits are designed to encourage personal wellness and smart healthcare decisions for you and your family while also knowing your life extends outside of work. Among our benefits, Humana provides medical, dental and vision benefits, 401(k) retirement savings plan, time off (including paid time off, company and personal holidays, volunteer time off, paid parental and caregiver leave), short-term and long-term disability, life insurance and many other opportunities. About Us About Inclusa: Inclusa manages the provision of a person-centered and community-focused approach to long-term care services and support to Family Care members across the state of Wisconsin. As a values-based organization devoted to building vibrant and inclusive communities, Inclusa deploys a unique approach to managed care with a trademarked model of support named Commonunity® which focuses on the belief in everyone, and from that belief, the common good for all is achieved. In 2022, Inclusa was acquired by Humana. This partnership will allow us to create a model of care that provides industry-leading support for members across the health care continuum. About Humana: Humana Inc. (NYSE: HUM) is a leading U.S. healthcare company. Through our Humana insurance services and our CenterWell healthcare services, we make it easier for the millions of people we serve to achieve their best health – delivering the care and service they need, when they need it. These efforts are leading to a better quality of life for people with Medicare and Medicaid, families, individuals, military service personnel, and communities at large. Learn more about what we offer at Humana.com and at CenterWell.com. Equal Opportunity Employer It is the policy of Humana not to discriminate against any employee or applicant for employment because of race, color, religion, sex, sexual orientation, gender identity, national origin, age, marital status, genetic information, disability or protected veteran status. It is also the policy of Humana to take affirmative action, in compliance with Section 503 of the Rehabilitation Act and VEVRAA, to employ and to advance in employment individuals with disability or protected veteran status, and to base all employment decisions only on valid job requirements. This policy shall apply to all employment actions, including but not limited to recruitment, hiring, upgrading, promotion, transfer, demotion, layoff, recall, termination, rates of pay or other forms of compensation and selection for training, including apprenticeship, at all levels of employment.

via universal intelligenceApply ›
View Lead Principal Engineer Product Quality and Reliability (f/m/div)
I

Lead Principal Engineer Product Quality and Reliability (f/m/div)

Infineon TechnologiesMunich, Bavaria, Germanyonsite

From $12/yr

#WeAreIn for jobs that impact everyone's life. Driving continuous improvement - ready to make quality a culture? As a Lead Principal Engineer Product Quality and Reliability on our Quality \& Excellence team uncover the root causes behind every single error, drive transformational improvements, and implement forward-thinking solutions to ensure our products exceed the highest standards of quality, reliability, and performance. Are you in? Your Role Key responsibilities in your new role * Be responsible for escalation and crisis management including on site customer communication and methodical problem-solving support * Be responsible for excursion management incl. defining, executing, assessing proper reliability tests and failure rate prediction * Collaborate with technical marketing and regional teams to foster requirement capturing by active customer engagement * Define risk and/or mission profile-based qualification and reliability testing concepts and assess test results during product development * Contribute to qualification assessments in development projects (Q-team) and act as Q-Team leader * Advise the team by providing risk assessment guidelines for risk material dispositions, reliability test referencing and product qualifications * Be a mentor, share knowledge and foster talents within the team * Transparently communicate within the international teams in Langen, Munich, Seoul and Cheonan Your Profile Qualifications And Skills To Help You Succeed As a Lead Principal Engineer Product Quality and Reliability you commit yourself to the results of your own team and the company as a whole. You are an integral part of innovations during product development. Furthermore, you are personally committed to the customer’s concerns and award them a high priority. Lastly, you set yourself ambitious goals. * M.Sc. degree in electronics, physics or similar, PhD is a plus * At least 12 years of strong experience in the automotive semiconductor or electronics business with direct customer contact * Broad product knowledge from wide band gap switch, gate driver, controller to (smart) IPM * Have a deep understanding of power semiconductor reliability tests and failure mechanisms * Solid knowledge on typical industrial and automotive applications (industrial drives, eCompressors, on board charger) * Extensive comprehension of industrial and automotive standards relevant to power semiconductors * Experience in crisis handling and with international customers * Ability and resilience to work under pressure * Proven technical leading * Excellent internal and customer communication skills * Proficiency in English, written and verbal, German or Korean are a plus * Availability to travel in Europe and Asia Contact: Darana Andrade #WeAreIn for driving decarbonization and digitalization. As a global leader in semiconductor solutions in power systems and IoT, Infineon enables game-changing solutions for green and efficient energy, clean and safe mobility, as well as smart and secure IoT. Together, we drive innovation and customer success, while caring for our people and empowering them to reach ambitious goals. Be a part of making life easier, safer and greener. Are you in? We are on a journey to create the best Infineon for everyone. This means we embrace diversity and inclusion and welcome everyone for who they are. At Infineon, we offer a working environment characterized by trust, openness, respect and tolerance and are committed to give all applicants and employees equal opportunities. We base our recruiting decisions on the applicant´s experience and skills. Learn more about our various contact channels. We look forward to receiving your resume, even if you do not entirely meet all the requirements of the job posting. Please let your recruiter know if they need to pay special attention to something in order to enable your participation in the interview process. Click here for more information about Diversity \& Inclusion at Infineon.

3 months agovia universal intelligenceApply ›
View Integration Reliability Engineer, Technical Operations
S

Integration Reliability Engineer, Technical Operations

StripeDublinonsite

€72,200 - €108,400/yr

About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world’s largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within

via universal intelligenceApply ›
View Integration Reliability Engineer, Technical Operations, Cards
S

Integration Reliability Engineer, Technical Operations, Cards

StripeToronto, Remote in Canadaremote

From CA$119,700/yr

About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world’s largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within

via universal intelligenceApply ›
View Senior Site Reliability Engineer (Auth0)
O

Senior Site Reliability Engineer (Auth0)

OktaBarcelona, Spainonsite

Get to know Okta Okta is The World’s Identity Company. We free everyone to safely use any technology, anywhere, on any device or app. Our flexible and neutral products, Okta Platform and Auth0 Platform, provide secure access, authentication, and automation, placing identity at the core of business security and growth. At Okta, we celebrate a variety of perspectives and experiences. We are not lookin

via universal intelligenceApply ›
View Senior Site Reliability Engineer, Compute
R

Senior Site Reliability Engineer, Compute

RobloxSan Mateo, CA, United Statesonsite

$195,780 - $242,100/yr

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people c

via universal intelligenceApply ›
View Staff Site Reliability Engineer-Federal, Security Clearance
Z

Staff Site Reliability Engineer-Federal, Security Clearance

ZscalerCrystal City, Virginia, USA

$119,000 - $170,000/yr

About Zscaler Zscaler is a pioneer and global leader in zero trust security. The world’s largest businesses, critical infrastructure organizations, and government agencies rely on Zscaler to secure users, branches, applications, data & devices, and to accelerate digital transformation initiatives. Distributed across more than 160 data centers globally, the Zscaler Zero Trust Exchange platform combined with advanced AI combats billions of

7 months agovia universal intelligenceApply ›
View Senior Database Reliability Engineer (DBRE)
O

Senior Database Reliability Engineer (DBRE)

OktaBellevue, Washingtononsite

$152,000 - $228,000/yr

Get to know Okta Okta is The World’s Identity Company. We free everyone to safely use any technology, anywhere, on any device or app. Our flexible and neutral products, Okta Platform and Auth0 Platform, provide secure access, authentication, and automation, placing identity at the core of business security and growth. At Okta, we celebrate a variety of perspectives and experiences. We are not lookin

via universal intelligenceApply ›
View Site Reliability Engineering Manager
O

Site Reliability Engineering Manager

OktaBarcelona, Spainonsite

Get to know Okta Okta is The World’s Identity Company. We free everyone to safely use any technology, anywhere, on any device or app. Our flexible and neutral products, Okta Platform and Auth0 Platform, provide secure access, authentication, and automation, placing identity at the core of business security and growth. At Okta, we celebrate a variety of perspectives and experiences. We are not lookin

via universal intelligenceApply ›
View Staff Site Reliability Engineer
O

Staff Site Reliability Engineer

OktaBengaluru, Indiaonsite

Get to know Okta Okta is The World’s Identity Company. We free everyone to safely use any technology, anywhere, on any device or app. Our flexible and neutral products, Okta Platform and Auth0 Platform, provide secure access, authentication, and automation, placing identity at the core of business security and growth. At Okta, we celebrate a variety of perspectives and experiences. We are not lookin

via universal intelligenceApply ›
View Staff Site Reliability Engineer-Federal, Security Clearance
Z

Staff Site Reliability Engineer-Federal, Security Clearance

ZscalerCrystal City, Virginia, USAonsite

$119,000 - $170,000/yr

About Zscaler Zscaler is a pioneer and global leader in zero trust security. The world’s largest businesses, critical infrastructure organizations, and government agencies rely on Zscaler to secure users, branches, applications, data & devices, and to accelerate digital transformation initiatives. Distributed across more than 160 data centers globally, the Zscaler Zero Trust Exchange platform combined with advanced AI combats billions of

via universal intelligenceApply ›
View Senior Reliability Engineer
E

Senior Reliability Engineer

Excelitas TechnologiesBillerica, MA, US

From $120,000/yr

Senior Reliability Engineer Location:Billerica, MA, US, 01821 Job Function: Engineering R\&D Job Country: United States Job City: Billerica Job Type: Regular Full Time ENABLE your future through light. Excelitas is a global technology leader with more than 7,500 employees, focused on delivering market-driven solutions to fulfill the illumination, optical, detection and imaging needs of OEMs and end-users across the biomedical, semiconductor, industrial, consumer products, scientific, security, defense and aerospace sectors. ENGAGE with us today and make your contribution to the future! Join the team that leading technology companies turn to for cutting-edge photonic innovation. At Excelitas Technologies you are how we EXCEL. Our facility in Billerica, Massachusetts is located in the metropolitan Boston area and specializes in the design and manufacturing of advanced photonics products based on MEMS and other micro-scale optical devices. Our unique technology enables the highest system performance in the most compact and cost-effective formats in a variety of applications. This site's Optical Coherence Tomography (OCT) products are laser-based imaging tools that have transformed ophthalmic care and made notable impacts on patients in cardiology, dermatology, oncology, gastroenterology, and other medical specialties. We are presently seeking Senior Reliability Engineer that will make vital contributions to new technology and product development for our growing range of cutting-edge optical engine products Main responsibilities: Work closely in a cross-functional teams to define, test, and validate laser system performance to meet the requirements of proposed products; Develop and execute on product qualification plans Develop, build test stations and equipment Manage and execute on failure analysis and related activities Troubleshoot Electronics and Optoelectronic Systems Support system prototyping Perform data analysis, compare numerical predictions and experimental results, present findings Test, characterization, and analysis of complex optical, opto-electronic systems to support new technology and product development; Optoelectronic system design; Mechanical design Manage Qual/Rel area and equipment Project and program management. Customer support Requirements: MSc in Electrical or Optical Engineering 5 plus years relevant industry experience; Direct experience with the design and implementation of optical interferometric systems, OCT, and/or swept-source lasers, or similar is desired; In-depth optical, opto-electronic design and laser engineering; Industrial experience in laser design, fiber optics Hands-on lab skills with optical equipment such as optical spectrum analyzers (OSAs), oscilloscopes, power meters, fiber optics, and data acquisition; Solid problem-solving methodology based on design of experiment, characterization and thorough data analysis; Self-motivated individual with ability to work independently and in a team environment. Manufacturing Engineer II Location: Billerica, MA, US, 01821 Job Function: Operations - Engineering Job Country: United States Job City: Billerica Job Type: Regular Full Time ENABLE your future through light. Excelitas is a global technology leader with more than 7,500 employees, focused on delivering market-driven solutions to fulfill the illumination, optical, detection and imaging needs of OEMs and end-users across the biomedical, semiconductor, industrial, consumer products, scientific, security, defense and aerospace sectors. ENGAGE with us today and make your contribution to the future! Join the team that leading technology companies turn to for cutting-edge photonic innovation. At Excelitas Technologies you are how we EXCEL. Excelitas is seeking leaders and innovators to join our global team! ENABLE your future through light. Excelitas is a global technology leader with more than 7,500 employees, focused on delivering market-driven solutions to fulfill the illumination, optical, detection and imaging needs of OEMs and end-users across the biomedical, semiconductor, industrial, consumer products, scientific, security, defense and aerospace sectors. ENGAGE with us today and make your contribution to the future! Join the team that leading technology companies turn to for cutting-edge photonic innovation. At Excelitas Technologies you are how we EXCEL. Our facility in Billerica, Massachusetts is located in the metropolitan Boston area and specializes in the design and manufacturing of advanced photonics products based on MEMS and other micro-scale optical devices. Our unique technology enables the highest system performance in the most compact and cost-effective formats in a variety of applications. This site's Optical Coherence Tomography (OCT) products are laser-based imaging tools that have transformed ophthalmic care and made notable impacts on patients in cardiology, dermatology, oncology, gastroenterology, and other medical specialties. We are presently seeking a Manufacturing Engineer that will make vital contributions to new technology and product development for our growing range of cutting-edge optical engine products. Main responsibilities: Plan and design complex manufacturing processes for our subcomponent and optical module assembly and test areas; Determine and specify the materials, equipment, and tools needed to achieve manufacturing goals and product specifications; Maximize efficiency by analyzing layout of equipment, workflow, assembly methods, and work force utilization; Write work instructions and train direct labor personnel on manufacturing processes; Monitor and improve yield and performance of modules and subcomponents; Analyze, troubleshoot, find root cause, correct manufacturing process issues, and establish appropriate process controls; Perform and support failure analysis of optical modules and subcomponents; Lead and report progress on continuous improvement projects and programs; Handles multiple projects and coordinates with other functions to ensure commitments are met and quality is maintained; Ensures a safe work environment by identifying/addressing near misses to prevent accidents; Support Operations to help meet production goals; Support transfer of new products from R\&D to manufacturing and harmonization with previously released products. Requirements: BS in Electrical or Optical Engineering or equivalent work experience; At least 5 years of manufacturing engineering experience; Hands-on skills with electro-optical equipment such as optical spectrum analyzers (OSAs), oscilloscopes, and power meters preferred; Familiarity with software-based testing is a plus. Experience with incorporating automation is a plus. Solid problem-solving skills using Six Sigma methodologies and thorough data analysis; Lean Manufacturing experience is a plus. Excellent communication and interpersonal skills, as well as strong organizational skills with superior attention to detail; Ability to be self-directed and work with minimal supervision; highly motivated self-starter able to work with multiple functions to address problems. Pay Range: $120,000 - $140,000 annually, depending on experience. This position requires the use of information which is subject to the International Traffic in Arms Regulations (ITAR) Visa sponsorship is not available for any position at Excelitas There is no relocation for this position. Equal Opportunity/Affirmative Action Employer Minorities/Females/Disability/Veteran/Gender Identity/Sexual Orientation Excelitas is seeking leaders and innovators to join our global team! #LI-AM1

recentlyvia universal intelligenceApply ›
View Reliability Engineering Manager
A

Reliability Engineering Manager

AmphenolYocumtown, PA, US

Position Summary: Reliability Engineering Manager Location: Vally Green, PA Amphenol High Speed Products Group is the market leader for high speed, high bandwidth electrical connectors for the Telecom/Datacom market (Mobile Networks, Storage, Servers, Routers, Switches, etc.). Our products help to enable the electronics revolution and remain a key enabler for all the major Tier 1 OEMs globally. We are currently seeking an experienced Reliability Engineering Manager to join our team. RESPONSIBILITIES: As the Reliability Engineering Manager you will drive product robustness, ensure compliance with customer and industry standards, and lead cross-functional efforts to predict, detect, and eliminate product failure risks across the product lifecycle. Guide team on developing test sequences and procedures to determine the long-term reliability of cable and connector assemblies that factor in customer use and application requirements Ensuring the accurate and timely completion of engineering tasks of engineers within the group as well as technicians in the lab Oversee and mentor engineering team; providing needed training and working with team members to grow within the organization Coordinating with peers and other engineering managers within the organization to share best practices and continuous improvement activities Lead the development, planning, and execution of reliability test plans for new product introduction and engineering changes. Define product qualification strategies aligned with application requirements and customer expectations. Drive failure mode analysis (FMEA), DfR (Design for Reliability), and HALT/HASS methods to assess and improve product robustness. Establish reliability metrics (FIT, DPPM, MTBF, escape rate projections) and track performance across the product portfolio. Collaborate closely with Design, Test Engineering, Quality, and Manufacturing teams to address reliability risks early in the design phase. Oversee execution of environmental, mechanical, and electrical stress testing (e.g., thermal cycling, vibration, humidity, etc.). Partner with suppliers and contract manufacturers to ensure compliance with reliability expectations and qualification requirements. Lead and manage root cause analysis of returned/failed units and implement corrective/preventive actions. Provide leadership, mentorship, and performance management for a team of reliability engineers and technicians. Prepare and present technical reports to internal and external stakeholders, including customers. QUALIFICATIONS: Bachelor’s degree in electrical engineering, Mechanical Engineering, Materials Science, or a related field. 5+ years of experience in reliability or product engineering, with at least 2+ years in a managerial or team lead role. Knowledge of reliability prediction models and field return data analysis. Experience with solid modeling tools - Pro/E and/or SolidWorks would be a plus Experience with data analysis software, MS Excel, Minitab would be a plus Experience developing test plans for mechanical and environmental stresses of high-speed electrical connectors; deep understanding of electrical connectors Experience specifying testing protocol for High-Speed products would be a plus Experience with Finite Element Analysis tools desirable, ANSYS experience would be a plus Ability to communicate with technical \& non-technical people Knowledge of copper alloys and high-performance engineering thermoplastics would be a plus Amphenol Corporation is proud of our reputation as an excellent employer. Our main focus is to provide the highest level of support and responsiveness to both our employees and our customers, the world's largest technology companies. Amphenol Corporation offers the opportunity for career growth within a global organization. We believe that Amphenol Corporation is unique in that every employee, regardless of his or her position, has the ability to positively impact the business. Amphenol is an “Equal Opportunity Employer” - Minority/Female/Disabled/Veteran/Sexual Orientation/Gender Identity/National Origin For additional company information please visit our website at https://www.amphenol-cs.com/

recentlyvia universal intelligenceApply ›
View Senior Site Reliability Engineer, Identity Platform
C

Senior Site Reliability Engineer, Identity Platform

CoinbaseRemote - USAremote

Ready to be pushed beyond what you think you’re capable of? At Coinbase, our mission is to increase economic freedom in the world. It’s a massive, ambitious opportunity that demands the best of us, every day, as we build the emerging onchain platform — and with it, the future global financial system. To achieve our mission, we’re seeking a very specific candidate. We want someone who is passionate about our mission and who believes in the power of cryp

7 months agovia universal intelligenceApply ›
View Site Reliability Engineer - Platform
C

Site Reliability Engineer - Platform

CoinbaseRemote - USAremote

Ready to be pushed beyond what you think you’re capable of? At Coinbase, our mission is to increase economic freedom in the world. It’s a massive, ambitious opportunity that demands the best of us, every day, as we build the emerging onchain platform — and with it, the future global financial system. To achieve our mission, we’re seeking a very specific candidate. We want someone who is passionate about our mission and who believes in the power of cryp

7 months agovia universal intelligenceApply ›
View Reliability Engineering Manager
A

Reliability Engineering Manager

AmphenolYocumtown, PA, US

Position Summary: Reliability Engineering Manager Location: Vally Green, PA Amphenol High Speed Products Group is the market leader for high speed, high bandwidth electrical connectors for the Telecom/Datacom market (Mobile Networks, Storage, Servers, Routers, Switches, etc.). Our products help to enable the electronics revolution and remain a key enabler for all the major Tier 1 OEMs globally. We are currently seeking an experienced Reliability Engineering Manager to join our team. RESPONSIBILITIES: As the Reliability Engineering Manager you will drive product robustness, ensure compliance with customer and industry standards, and lead cross-functional efforts to predict, detect, and eliminate product failure risks across the product lifecycle. Guide team on developing test sequences and procedures to determine the long-term reliability of cable and connector assemblies that factor in customer use and application requirements Ensuring the accurate and timely completion of engineering tasks of engineers within the group as well as technicians in the lab Oversee and mentor engineering team; providing needed training and working with team members to grow within the organization Coordinating with peers and other engineering managers within the organization to share best practices and continuous improvement activities Lead the development, planning, and execution of reliability test plans for new product introduction and engineering changes. Define product qualification strategies aligned with application requirements and customer expectations. Drive failure mode analysis (FMEA), DfR (Design for Reliability), and HALT/HASS methods to assess and improve product robustness. Establish reliability metrics (FIT, DPPM, MTBF, escape rate projections) and track performance across the product portfolio. Collaborate closely with Design, Test Engineering, Quality, and Manufacturing teams to address reliability risks early in the design phase. Oversee execution of environmental, mechanical, and electrical stress testing (e.g., thermal cycling, vibration, humidity, etc.). Partner with suppliers and contract manufacturers to ensure compliance with reliability expectations and qualification requirements. Lead and manage root cause analysis of returned/failed units and implement corrective/preventive actions. Provide leadership, mentorship, and performance management for a team of reliability engineers and technicians. Prepare and present technical reports to internal and external stakeholders, including customers. QUALIFICATIONS: Bachelor’s degree in electrical engineering, Mechanical Engineering, Materials Science, or a related field. 5+ years of experience in reliability or product engineering, with at least 2+ years in a managerial or team lead role. Knowledge of reliability prediction models and field return data analysis. Experience with solid modeling tools - Pro/E and/or SolidWorks would be a plus Experience with data analysis software, MS Excel, Minitab would be a plus Experience developing test plans for mechanical and environmental stresses of high-speed electrical connectors; deep understanding of electrical connectors Experience specifying testing protocol for High-Speed products would be a plus Experience with Finite Element Analysis tools desirable, ANSYS experience would be a plus Ability to communicate with technical \& non-technical people Knowledge of copper alloys and high-performance engineering thermoplastics would be a plus Amphenol Corporation is proud of our reputation as an excellent employer. Our main focus is to provide the highest level of support and responsiveness to both our employees and our customers, the world's largest technology companies. Amphenol Corporation offers the opportunity for career growth within a global organization. We believe that Amphenol Corporation is unique in that every employee, regardless of his or her position, has the ability to positively impact the business. Amphenol is an “Equal Opportunity Employer” - Minority/Female/Disabled/Veteran/Sexual Orientation/Gender Identity/National Origin For additional company information please visit our website at https://www.amphenol-cs.com/

recentlyvia universal intelligenceApply ›
View Director- Site Reliability Engineer
O

Director- Site Reliability Engineer

OktaBengaluru, Indiaonsite

Get to know Okta Okta is The World’s Identity Company. We free everyone to safely use any technology, anywhere, on any device or app. Our flexible and neutral products, Okta Platform and Auth0 Platform, provide secure access, authentication, and automation, placing identity at the core of business security and growth. At Okta, we celebrate a variety of perspectives and experiences. We are not lookin

via universal intelligenceApply ›
View Staff Site Reliability Engineer, Cloud Efficiency
O

Staff Site Reliability Engineer, Cloud Efficiency

OktaBengaluru, India

Get to know OktaOkta is The World’s Identity Company. We free everyone to safely use any technology, anywhere, on any device or app. Our flexible and neutral products, Okta Platform and Auth0 Platform, provide secure access, authentication, and automation, placing identity at the core of business security and growth.At Okta, we celebrate a variety of perspectives and experiences. We are not lookin

7 months agovia universal intelligenceApply ›
View Staff Site Reliability Engineer - Federal
Z

Staff Site Reliability Engineer - Federal

ZscalerCrystal City, Virginia, USA; Remote - Virginia, USAremote

About Zscaler Zscaler is a pioneer and global leader in zero trust security. The world’s largest businesses, critical infrastructure organizations, and government agencies rely on Zscaler to secure users, branches, applications, data & devices, and to accelerate digital transformation initiatives. Distributed across more than 160 data centers globally, the Zscaler Zero Trust Exchange platform combined with advanced AI combats billions of

7 months agovia universal intelligenceApply ›
View Software Engineer, Reliability
R

Software Engineer, Reliability

RobloxSan Mateo, CA, United States

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people c

7 months agovia universal intelligenceApply ›
View Senior Manager of Site Reliability Engineering
G

Senior Manager of Site Reliability Engineering

GoDaddyArizona (hybrid role, not eligible in Alaska, Mississippi, North Dakota, or the Virgin Islands; also not considering candidates from California, Seattle, or NYC)remote

From $401,000/yr

Location Details: Arizona At GoDaddy the future of work looks different for each team. Some teams work in the office full-time, others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely. This is a hybrid position. You’ll divide your time between working remotely from your home and an office, so you should live within commuting distance. Hybrid teams may work in-office as much as a few times a week or as little as once a month or quarter, as decided by leadership. The hiring manager can share more about what hybrid work might look like for this team. This position is not eligible to be performed in Alaska, Mississippi, North Dakota, or the Virgin Islands. GoDaddy is not currently considering candidates for this role in California, Seattle, or NYC. Join our team… GoDaddy is seeking an exceptional Senior Manager of Site Reliability Engineering to lead our Public Key Infrastructure (PKI) operations. In this pivotal role, you

recentlyvia universal intelligenceApply ›
View Senior Manager, Site Reliability Engineering - Infrastructure Platform
O

Senior Manager, Site Reliability Engineering - Infrastructure Platform

OktaBellevue, Washington

Get to know OktaOkta is The World’s Identity Company. We free everyone to safely use any technology, anywhere, on any device or app. Our flexible and neutral products, Okta Platform and Auth0 Platform, provide secure access, authentication, and automation, placing identity at the core of business security and growth.At Okta, we celebrate a variety of perspectives and experiences. We are not lookin

7 months agovia universal intelligenceApply ›
View Engineering Manager, Database Reliability, Scalability & Operations
G

Engineering Manager, Database Reliability, Scalability & Operations

GitLabNot specified

From $100,000/yr

Engineering Manager, Database Reliability, Scalability & Operations GitLab is an open-core software company that develops the most comprehensive AI-powered DevSecOps Platform, used by more than 100,000 organizations. Our mission is to enable everyone to contribute to and co-create the software that powers our world. When everyone can contribute, consumers become contributors, significantly accelerating human progress. Our platform unites teams and organizations, breaking down barriers and redefining what's possible in software development. Thanks to products like Duo Enterprise and Duo Agent Platform, customers get AI benefits at every stage of the SDLC. The same principles built into our products are reflected in how our team works: we embrace AI as a core productivity multiplier, with all team members expected to incorporate AI into their daily workflows to drive efficiency, innovation, and impact. GitLab is where careers accelerate, innovation flourishes, and every voice is valued.

Experience leading distribu...Background in hiring, coach...Applied experience designin...Ability to define database ...Practice setting clear meas...+2 more
recentlyvia universal intelligenceApply ›
View Staff Site Reliability Engineer-Federal, Security Clearance
Z

Staff Site Reliability Engineer-Federal, Security Clearance

ZscalerCrystal City, Virginia, USA

About Zscaler Zscaler is a pioneer and global leader in zero trust security. The world’s largest businesses, critical infrastructure organizations, and government agencies rely on Zscaler to secure users, branches, applications, data & devices, and to accelerate digital transformation initiatives. Distributed across more than 160 data centers globally, the Zscaler Zero Trust Exchange platform combined with advanced AI combats billions of

7 months agovia universal intelligenceApply ›
View Senior Site Reliability Engineer, Compute
R

Senior Site Reliability Engineer, Compute

RobloxSan Mateo, CA, United States

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people c

7 months agovia universal intelligenceApply ›
View Senior Network Reliability Engineer
R

Senior Network Reliability Engineer

RobloxSan Mateo, CA, United States

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people c

7 months agovia universal intelligenceApply ›
View Director- Site Reliability Engineer
O

Director- Site Reliability Engineer

OktaBengaluru, India

Get to know OktaOkta is The World’s Identity Company. We free everyone to safely use any technology, anywhere, on any device or app. Our flexible and neutral products, Okta Platform and Auth0 Platform, provide secure access, authentication, and automation, placing identity at the core of business security and growth.At Okta, we celebrate a variety of perspectives and experiences. We are not lookin

7 months agovia universal intelligenceApply ›
View Staff Site Reliability Engineer, Streaming
A

Staff Site Reliability Engineer, Streaming

AlpacaRemoteremote

From $24/yr

Staff Site Reliability Engineer, Streaming Who We Are: Alpaca is a US-headquartered self-clearing broker-dealer and brokerage infrastructure for stocks, ETFs, options, crypto, fixed income, 24/5 trading, and more. Our recent Series C funding round brought our total investment to over $170 million, fueling our ambitious vision. Amongst our subsidiaries, Alpaca is a licensed financial services company, serving hundreds of financial institutions across 40 countries with our institutional-grade APIs. This includes broker-dealers, investment advisors, wealth managers, hedge funds, and crypto exchanges, totalling over 6 million brokerage accounts. Our global team is a diverse group of experienced engineers, traders, and brokerage professionals who are working to achieve our mission of opening financial services to everyone on the planet. We're deeply committed to open-source contributions and fostering a vibrant community, continuously enhancing our award-winning, developer-friendly API and

Site Reliability EngineeringPerformance Engineering
recentlyvia universal intelligenceApply ›
View Algorithmic Trading, Senior Site Reliability Engineer (SRE)
B

Algorithmic Trading, Senior Site Reliability Engineer (SRE)

BTIGNew York, New York, United States

Create a Job Alert Level-up your career by having opportunities at BTIG sent directly to your inbox. Job Algorithmic Trading, Senior Site Reliability Engineer (SRE) New York, New York, United States Franchise Sales, Healthcare, Vice President/Director 2026 Full-Time Investment Banking Analyst San Francisco, California, United States Investment Banking, Senior Analyst/Associate, DCM Equity Research Associate, Medical Technology Chicago, New York Equity Research Senior Associate, Software New York, San Francisco ; San Francisco, California, United States Equity Research Supervisory Analyst Research, Equity Research Associate, Biotechnology Structured Products, Analyst Structured Products, Credit Sales Executive, Associate/Vice President Structured Products, Mortgage Origination Sales, Vice President Technology, Service Desk Associate Technology, Sr. PostgreSQL Database Engineer, Vice President Technology, Trading Floor Desktop Support Engineer

recentlyvia universal intelligenceApply ›
View Senior Site Reliability Engineer
R

Senior Site Reliability Engineer

RegrelloSeattle, San Francisco, New York or remote in US, Canada, Mexicoremote

From $40/yr

As a Senior Site Reliability Engineer at Regrello, you'll shape the developer platform, collaborate with customers, and ensure the reliability and security of infrastructure and applications. Company Regrello is a 40-person startup reimagining automation in supply chains, in which companies still communicate about $13T of annual shipments almost entirely via email. This is a $220-billion, Amazon-sized market opportunity that’s been largely overlooked for over 20 years. Our team has experience building billion-dollar companies and includes veterans from Oracle, Palantir, Apple, Amazon, and Microsoft. Our customers include some of the largest electronics manufacturers in the world, all of which are hungry for us to succeed. Regrello has been funded by Andreessen Horowitz, Tiger Global, Dell, Bloomberg, and Ram Sriram (first investor in Google). Product We are building a global operating network that finally enables supply-chain companies to collaborate within one platform. Our workflow e

AWSAzureCircleCIGCPGithub Actions+4 more
recentlyvia universal intelligenceApply ›
View Staff Site Reliability Engineer (SRE & Platform Reliability)
A

Staff Site Reliability Engineer (SRE & Platform Reliability)

AffirmRemote Polandremote

Affirm is reinventing credit to make it more honest and friendly, giving consumers the flexibility to buy now and pay later without any hidden fees or compounding interest.Affirm is reinventing credit to make it more honest and friendly, giving consumers the flexibility to buy now and pay later without any hidden fees or compounding interest. Site Reliability Engineering at Affirm is a small, yet crucial, team that helps ou

7 months agovia universal intelligenceApply ›
View Senior Machine Learning Site Reliability Engineer
P

Senior Machine Learning Site Reliability Engineer

PrimaMilan / Madrid / Londonremote

From $201/yr

Senior Machine Learning Site Reliability Engineer Milan / Madrid / London Engineering – Engineering / Permanent Employment / Remote Are you looking for a new challenge? Fancy helping us shape the future of motor insurance? Prima could be the place for you. Since 2015, we’ve been using our love of data and tech to rethink motor insurance and bring drivers a great experience at a great price. Our story began in Italy, where we’ve quickly become the number one online motor insurance provider. In fact, we’re trusted by over 5 million drivers. And now we’re expanding to help millions more drivers in the UK and Spain. To help fuel that growth, we need a Senior Machine Learning Site Reliability Engineer to join our Infrastructure team. This team is the beating heart of Prima. You’ll be joining over 300 engineers across software development, infrastructure, operations and security. Fueled by curiosity, experimentation and collaboration, you’ll help deliver scalable, impactful solutions that sh

SRE practices in productionAWS expertiseKubernetesnetworkingDNS+9 more
recentlyvia universal intelligenceApply ›
View Tech Lead, Platform Reliability Engineering
A

Tech Lead, Platform Reliability Engineering

AsanaWarsaw

We're looking for a Tech Lead to own technical direction for Asana's Platform Reliability Engineering team. This role is for someone who thrives as a technical leader without needing to take on people management. At Asana, Tech Leads own technical execution, architecture, and long-term engineering strategy. You'll lead complex projects, make key technical decisions, and set the standard for engineering quality. You'll guide ICs through technical design and project leadership, and partn

7 months agovia universal intelligenceApply ›
View Senior Site Reliability Engineer (SRE & Platform Reliability
A

Senior Site Reliability Engineer (SRE & Platform Reliability

AffirmRemote Polandremote

Affirm is reinventing credit to make it more honest and friendly, giving consumers the flexibility to buy now and pay later without any hidden fees or compounding interest.Site Reliability Engineering at Affirm is a small, yet crucial, team that helps our Engineering partners to “Operate What They Own” with excellence to protect their customers’ experience. SRE accomplishes this through defining frameworks and best practices for operating applications,

7 months agovia universal intelligenceApply ›
View Site Reliability Engineer
G

Site Reliability Engineer

GoDaddyIndia, remoteremote

From $401,000/yr

Location Details: India, remote At GoDaddy, the future of work looks different for each team. Some teams work in the office full-time, others have a hybrid arrangement (they work remotely some days and in the office some days) , and some work entirely remotely. This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings. Join Our Team GoDaddy is seeking a Site Reliability Engineer to join our team and play a crucial role in building and maintaining our cutting-edge internal tools. These tools enable thousands of developers to deploy seamlessly to production and efficiently manage their AWS infrastructure. As a Site Reliability Engineer (SRE), you will collaborate with multinational development teams to ensure the smooth deployment and ongoing maintenance of our platform. What you'll get to do... Your experience should include... We've got your back... We offer a range of total rewards

recentlyvia universal intelligenceApply ›
View Site Reliability Engineer Intern (Summer 2026)
O

Site Reliability Engineer Intern (Summer 2026)

OktaBellevue, Washington

Get to know OktaOkta is The World’s Identity Company. We free everyone to safely use any technology, anywhere, on any device or app. Our flexible and neutral products, Okta Platform and Auth0 Platform, provide secure access, authentication, and automation, placing identity at the core of business security and growth.At Okta, we celebrate a variety of perspectives and experiences. We are not lookin

7 months agovia universal intelligenceApply ›
View Principal Site Reliability Engineer
O

Principal Site Reliability Engineer

OktaBengaluru, India

Get to know OktaOkta is The World’s Identity Company. We free everyone to safely use any technology, anywhere, on any device or app. Our flexible and neutral products, Okta Platform and Auth0 Platform, provide secure access, authentication, and automation, placing identity at the core of business security and growth.At Okta, we celebrate a variety of perspectives and experiences. We are not lookin

7 months agovia universal intelligenceApply ›
View Senior Site Reliability Engineer - Infrastructure
U

Senior Site Reliability Engineer - Infrastructure

UnderdogfantasyUnited States/Remoteremote

At Underdog, we make sports more fun. Our thesis is simple: build the best products and we’ll build the biggest company in the space, because there’s so much more to be built for sports fans. We’re just over five years in, and we’re one of the fastest-growing sports companies ever, most recently valued at $1.3B. And it’s still the early days. We’ve built and scaled multiple games and products across fantasy sports, sports betting, and prediction market

7 months agovia universal intelligenceApply ›
View Senior Site Reliability Engineer
A

Senior Site Reliability Engineer

AffirmRemote Spainremote

Affirm is reinventing credit to make it more honest and friendly, giving consumers the flexibility to buy now and pay later without any hidden fees or compounding interest.Affirm is looking for a Senior Software Engineer for our Cloud Compute team to play a pivotal role in ensuring the robust and scalable foundation of our entire platform. As a fully remote team based in Spain, we are the engine room that manages all of Affirm's Kubernetes clusters. Our

7 months agovia universal intelligenceApply ›
View Senior Site Reliability Engineer (Auth0)
O

Senior Site Reliability Engineer (Auth0)

OktaBarcelona, Spain

Get to know OktaOkta is The World’s Identity Company. We free everyone to safely use any technology, anywhere, on any device or app. Our flexible and neutral products, Okta Platform and Auth0 Platform, provide secure access, authentication, and automation, placing identity at the core of business security and growth.At Okta, we celebrate a variety of perspectives and experiences. We are not lookin

7 months agovia universal intelligenceApply ›
View Software Engineer, Network Performance & Reliability
C

Software Engineer, Network Performance & Reliability

CloudflareHybridhybrid

About Us At Cloudflare, we are on a mission to help build a better Internet. Today the company runs one of the world’s largest networks that powers millions of websites and other Internet properties for customers ranging from individual bloggers to SMBs to Fortune 500 companies. Cloudflare protects and accelerates any Internet application online without adding hardware, installing software, or changing a line of code. Internet prope

7 months agovia universal intelligenceApply ›
View Staff Site Reliability Engineer
O

Staff Site Reliability Engineer

OktaBengaluru, India

Get to know OktaOkta is The World’s Identity Company. We free everyone to safely use any technology, anywhere, on any device or app. Our flexible and neutral products, Okta Platform and Auth0 Platform, provide secure access, authentication, and automation, placing identity at the core of business security and growth.At Okta, we celebrate a variety of perspectives and experiences. We are not lookin

7 months agovia universal intelligenceApply ›
View Senior Site Reliability Engineer (SRE & Platform Reliability)
A

Senior Site Reliability Engineer (SRE & Platform Reliability)

AffirmRemote Spainremote

Affirm is reinventing credit to make it more honest and friendly, giving consumers the flexibility to buy now and pay later without any hidden fees or compounding interest.Site Reliability Engineering at Affirm is a small, yet crucial, team that helps our Engineering partners to “Operate What They Own” with excellence to protect their customers’ experience. SRE accomplishes this through defining frameworks and best practices for operating applications,

7 months agovia universal intelligenceApply ›

Showing 50 of 94 jobs. Sign up to see all results and set up alerts.

Companies hiring for Reliability Engineer

Reliability Engineer Salary Data

View detailed salary ranges from 21 listings with reported compensation ›

Track these jobs with Jobfu

Get AI-powered resume tailoring, application tracking, and job alerts for Reliability Engineer roles. Sign up free.

Sign up free