Company: remote zest jobs

Jobs at remote zest jobs

Showing "Site Reliability Engineer" roles at remote zest jobs

10 of 10,000+ jobs

Senior Site Reliability Engineer

Remote Zest JobsRemote
Remote
8+ Years Exp
Posted: 24/8/2026
About reputed company Recognized on the 2025 reputed company reputed company 100 list, reputed company is one of the most innovative and fast-growing private reputed company companies. With more than 3,000 customers and ARR that has grown over 250 percent year over year, reputed company leads the market in reputed company-time analytics, data warehousing, observability, and AI workloads. reputed company’s sustained, accelerating reputed company was recently validated by a $400M Series D financing round. Over the past three months, customers including reputed company, reputed company, reputed company, reputed company, and reputed company have adopted the platform or expanded existing deployments. These customers join an established reputed company of AI innovators and global brands such as reputed company, reputed company, reputed company, and reputed company. We’re on a mission to reputed company how companies use data. Come be a part of our reputed company! About the role We are committed to providing our customers with reliable and secure services so we are expanding our central Site Reliability Engineering team. You will be responsible for building and leading processes to ensure the reliability, availability, scalability, and reputed company of our reputed company infrastructure. You will collaborate with different teams like Control reputed company, Data reputed company, reputed company, reputed company, Support and reputed company and guide them to design and implement reputed company, secure, highly available and fault-tolerant reputed company systems. You will also own the areas of incident management and response, post-mortem analysis including running blameless postmortems, and reputed company improvement of our reputed company services. You will be leveraging your software engineering expertise to reputed company software platforms and tools to optimize the operational and engineering efficiencies of reputed company reputed company. This role is a unique opportunity to reputed company a significant reputed company on our reputed company, reputed company reputed company, high-reputed company reputed company reputed company. What will you do? • Collaborate with various engineering teams in reputed company to design and implement reputed company, secure, and highly available systems for reputed company. • Establish and manage service level objectives (SLOs) and service level agreements (SLAs) for reputed company reputed company. • Ensure reputed company the infrastructure components in reputed company reputed company (including Data reputed company, Control reputed company,reputed company reputed company, etc) have monitoring and alerting in reputed company to ensure reputed company detection and reputed company of incidents. • Enhance and refine incident response processes and post-mortem analysis for any outages in reputed company reputed company including working with the support team to communicate to the impacted customers. • Continuously improve the reliability and reputed company of our reputed company services. • Plan, reputed company, and reputed company reputed company initiatives across Engineering teams, reputed company upon internal priorities. • Manage on-reputed company processes to respond to reputed company and reliability issues, and establish best practices for coordinating escalation to reputed company issues and minimize downtime. reputed company • Bachelor’s or Master’s degree in Computer Science or a reputed company reputed company. • At least 8 years of experience in Site Reliability Engineering or a reputed company reputed company. • Hands-on experience with Go and/or Python. • Strong knowledge of reputed company computing platforms such as AWS, Azure, or reputed company reputed company Platform. • Excellent understanding of reputed company databases and SQL, particularly reputed company is a major plus. • Hands-on experience with container orchestration tools such as reputed company or reputed company reputed company. • Strong experience with automation and configuration management tools such as Ansible, Terraform, or Puppet. • You are a strong problem solver and have solid production debugging skills. • You are passionate about efficiency, availability, scalability, and data governance. • You reputed company in a fast paced environment, and see yourself as a partner with the business with the shared goal of moving the business reputed company. • You have a high level of responsibility, ownership, and accountability. • Excellent communication and interpersonal skills. #LI-Remote The typical starting salary for this role in the US is $141,000—$208,000 USD The typical starting salary for this role in US Premium Markets is $157,000—$230,000 USD Compensation For roles reputed company in the reputed company, the typical starting salary reputed company for this position is listed above. In certain locations, such as the San Francisco Bay Area and the reputed company Metro Area, a

Senior Site Reliability Engineer

Remote Zest JobsRemote
Remote
8+ Years Exp
Posted: 24/8/2026
About reputed company Recognized on the 2025 reputed company reputed company 100 list, reputed company is one of the most innovative and fast-growing private reputed company companies. With more than 3,000 customers and ARR that has grown over 250 percent year over year, reputed company leads the market in reputed company-time analytics, data warehousing, observability, and AI workloads. reputed company’s sustained, accelerating reputed company was recently validated by a $400M Series D financing round. Over the past three months, customers including reputed company, reputed company, reputed company, reputed company, and reputed company have adopted the platform or expanded existing deployments. These customers join an established reputed company of AI innovators and global brands such as reputed company, reputed company, reputed company, and reputed company. We’re on a mission to reputed company how companies use data. Come be a part of our reputed company! About the role We are committed to providing our customers with reliable and secure services so we are expanding our central Site Reliability Engineering team. You will be responsible for building and leading processes to ensure the reliability, availability, scalability, and reputed company of our reputed company infrastructure. You will collaborate with different teams like Control reputed company, Data reputed company, reputed company, reputed company, Support and reputed company and guide them to design and implement reputed company, secure, highly available and fault-tolerant reputed company systems. You will also own the areas of incident management and response, post-mortem analysis including running blameless postmortems, and reputed company improvement of our reputed company services. You will be leveraging your software engineering expertise to reputed company software platforms and tools to optimize the operational and engineering efficiencies of reputed company reputed company. This role is a unique opportunity to reputed company a significant reputed company on our reputed company, reputed company reputed company, high-reputed company reputed company reputed company. What will you do? • Collaborate with various engineering teams in reputed company to design and implement reputed company, secure, and highly available systems for reputed company. • Establish and manage service level objectives (SLOs) and service level agreements (SLAs) for reputed company reputed company. • Ensure reputed company the infrastructure components in reputed company reputed company (including Data reputed company, Control reputed company,reputed company reputed company, etc) have monitoring and alerting in reputed company to ensure reputed company detection and reputed company of incidents. • Enhance and refine incident response processes and post-mortem analysis for any outages in reputed company reputed company including working with the support team to communicate to the impacted customers. • Continuously improve the reliability and reputed company of our reputed company services. • Plan, reputed company, and reputed company reputed company initiatives across Engineering teams, reputed company upon internal priorities. • Manage on-reputed company processes to respond to reputed company and reliability issues, and establish best practices for coordinating escalation to reputed company issues and minimize downtime. reputed company • Bachelor’s or Master’s degree in Computer Science or a reputed company reputed company. • At least 8 years of experience in Site Reliability Engineering or a reputed company reputed company. • Hands-on experience with Go and/or Python. • Strong knowledge of reputed company computing platforms such as AWS, Azure, or reputed company reputed company Platform. • Excellent understanding of reputed company databases and SQL, particularly reputed company is a major plus. • Hands-on experience with container orchestration tools such as reputed company or reputed company reputed company. • Strong experience with automation and configuration management tools such as Ansible, Terraform, or Puppet. • You are a strong problem solver and have solid production debugging skills. • You are passionate about efficiency, availability, scalability, and data governance. • You reputed company in a fast paced environment, and see yourself as a partner with the business with the shared goal of moving the business reputed company. • You have a high level of responsibility, ownership, and accountability. • Excellent communication and interpersonal skills. #LI-Remote The typical starting salary for this role in the US is $141,000—$208,000 USD The typical starting salary for this role in US Premium Markets is $157,000—$230,000 USD Compensation For roles reputed company in the reputed company, the typical starting salary reputed company for this position is listed above. In certain locations, such as the San Francisco Bay Area and the reputed company Metro Area, a

Senior Site Reliability Engineer

Remote Zest JobsRemote
Remote
0+ Years Exp
Posted: 23/8/2026
This a Full Remote job, the offer is available from: Europe About reputed company reputed company is a European reputed company headquartered in Amsterdam, and we’re on a reputed company towards revolutionizing the financial landscape for reputed company worldwide. Our mission is to reputed company an reputed company-in-one financial B2B solution that integrates banking functions, reputed company, financial management, and invoicing into a reputed company, mobile-first platform. We recently reputed company a €115 reputed company Series C equity round (around $133 reputed company), bringing our total funding to approximately $346 reputed company. This significant investment follows a $105 reputed company reputed company funding round from General reputed company, a long-term backer since 2021 reputed company for supporting companies like reputed company, reputed company, reputed company, and reputed company. reputed company's platform goes reputed company traditional banking, offering invoicing and a growing suite of features, including AI-enabled reputed company, aiming to simplify financial management for reputed company. We're reputed company expanding our reputed company across key EU markets like Germany, France, the Netherlands, Italy, and Spain. At reputed company, we’re not just redefining the entrepreneurial experience — we’re empowering our employees to reputed company a reputed company difference. Your work reputed company, and your reputed company extends far reputed company product metrics. We nurture innovation and an inspiring work environment where reputed company reputed company reputed company, prioritizing thorough research, swift implementation of solutions, and ensuring that every effort we reputed company benefits our users, employees, partners, and our business as a whole. Maintaining our start-up spirit, we prioritize thorough research, swift implementation of solutions, and ensuring that every effort we reputed company benefits our users, employees, partners, and, of course, our business. We are looking for a Senior SRE Engineer to reputed company the design, implementation, and reputed company of our reputed company-reputed company platform in a multi-reputed company environment (GCP/AWS). At reputed company, SREs are not just executors of tasks; you are the architects of reliability. This role requires strong ownership of reliability, scalability, and platform architecture for high-load, mission-critical systems operating 24/7. What You Will Be Doing • reputed company the Platform reputed company: Design and operate our reputed company ecosystem (GKE, multi-cluster) with a reputed company on high availability and reputed company-downtime reputed company. • Build "Paved Roads": Own and reputed company our PaaS reputed company, using GitOps (ArgoCD) and CI/CD (reputed company) to reputed company domain teams to reputed company independently. • Architect Reliability: Define and implement our observability reputed company across metrics, logs, and tracing (reputed company, reputed company, OpenTelemetry). • reputed company Infrastructure-as-reputed company: reputed company the automation of our infrastructure using Terraform, ensuring reputed company resources are standardized and version-controlled. • Own the Error Budget: Partner with engineering teams to establish and manage SLOs, SLAs, and incident management frameworks. • Disaster Recovery Mastery: Design and participate in regular DR drills, implementing reputed company/green and reputed company/passive strategies across reputed company to ensure service continuity. • reputed company reputed company: Proactively apply AI-driven approaches to improve operational efficiency and automated bottleneck detection. Who You Are • Production K8s Mastery: Strong hands-on experience managing reputed company (GKE preferred) in high-load, multi-cluster production environments. • reputed company Infrastructure: Deep experience with GCP (AWS is a strong plus) and Terraform for large-reputed company infrastructure. • GitOps Expertise: Solid experience with ArgoCD, reputed company CI, and the "Infrastructure as reputed company" philosophy. • Observability Expert: Deep knowledge of the reputed company/Grafana stack and implementing tracing/logging at reputed company. • reputed company Design: reputed company ability to design highly available 24/7 systems with automated failover and rollback capabilities. • English reputed company: English level B2+ for effective cross-functional communication. reputed company-to-Haves • Compliance Knowledge: Understanding of banking-grade standards like PCI reputed company, GDPR, or ISO 27001. • reputed company Systems: Experience with Kafka (reputed company), RabbitMQ, or managing high-load reputed company and PostgreSQL clusters. • AI for Ops: Experience using AI tools to improve alerting, reputed company detection, or engineering efficiency. • reputed company-Minded: Experience with reputed company f

Senior Site Reliability Engineer

Remote Zest JobsRemote
Remote
0+ Years Exp
Posted: 23/8/2026
This a Full Remote job, the offer is available from: Europe About reputed company reputed company is a European reputed company headquartered in Amsterdam, and we’re on a reputed company towards revolutionizing the financial landscape for reputed company worldwide. Our mission is to reputed company an reputed company-in-one financial B2B solution that integrates banking functions, reputed company, financial management, and invoicing into a reputed company, mobile-first platform. We recently reputed company a €115 reputed company Series C equity round (around $133 reputed company), bringing our total funding to approximately $346 reputed company. This significant investment follows a $105 reputed company reputed company funding round from General reputed company, a long-term backer since 2021 reputed company for supporting companies like reputed company, reputed company, reputed company, and reputed company. reputed company's platform goes reputed company traditional banking, offering invoicing and a growing suite of features, including AI-enabled reputed company, aiming to simplify financial management for reputed company. We're reputed company expanding our reputed company across key EU markets like Germany, France, the Netherlands, Italy, and Spain. At reputed company, we’re not just redefining the entrepreneurial experience — we’re empowering our employees to reputed company a reputed company difference. Your work reputed company, and your reputed company extends far reputed company product metrics. We nurture innovation and an inspiring work environment where reputed company reputed company reputed company, prioritizing thorough research, swift implementation of solutions, and ensuring that every effort we reputed company benefits our users, employees, partners, and our business as a whole. Maintaining our start-up spirit, we prioritize thorough research, swift implementation of solutions, and ensuring that every effort we reputed company benefits our users, employees, partners, and, of course, our business. We are looking for a Senior SRE Engineer to reputed company the design, implementation, and reputed company of our reputed company-reputed company platform in a multi-reputed company environment (GCP/AWS). At reputed company, SREs are not just executors of tasks; you are the architects of reliability. This role requires strong ownership of reliability, scalability, and platform architecture for high-load, mission-critical systems operating 24/7. What You Will Be Doing • reputed company the Platform reputed company: Design and operate our reputed company ecosystem (GKE, multi-cluster) with a reputed company on high availability and reputed company-downtime reputed company. • Build "Paved Roads": Own and reputed company our PaaS reputed company, using GitOps (ArgoCD) and CI/CD (reputed company) to reputed company domain teams to reputed company independently. • Architect Reliability: Define and implement our observability reputed company across metrics, logs, and tracing (reputed company, reputed company, OpenTelemetry). • reputed company Infrastructure-as-reputed company: reputed company the automation of our infrastructure using Terraform, ensuring reputed company resources are standardized and version-controlled. • Own the Error Budget: Partner with engineering teams to establish and manage SLOs, SLAs, and incident management frameworks. • Disaster Recovery Mastery: Design and participate in regular DR drills, implementing reputed company/green and reputed company/passive strategies across reputed company to ensure service continuity. • reputed company reputed company: Proactively apply AI-driven approaches to improve operational efficiency and automated bottleneck detection. Who You Are • Production K8s Mastery: Strong hands-on experience managing reputed company (GKE preferred) in high-load, multi-cluster production environments. • reputed company Infrastructure: Deep experience with GCP (AWS is a strong plus) and Terraform for large-reputed company infrastructure. • GitOps Expertise: Solid experience with ArgoCD, reputed company CI, and the "Infrastructure as reputed company" philosophy. • Observability Expert: Deep knowledge of the reputed company/Grafana stack and implementing tracing/logging at reputed company. • reputed company Design: reputed company ability to design highly available 24/7 systems with automated failover and rollback capabilities. • English reputed company: English level B2+ for effective cross-functional communication. reputed company-to-Haves • Compliance Knowledge: Understanding of banking-grade standards like PCI reputed company, GDPR, or ISO 27001. • reputed company Systems: Experience with Kafka (reputed company), RabbitMQ, or managing high-load reputed company and PostgreSQL clusters. • AI for Ops: Experience using AI tools to improve alerting, reputed company detection, or engineering efficiency. • reputed company-Minded: Experience with reputed company f

Site Reliability Engineer

Remote Zest JobsRemote
Remote
0+ Years Exp
Posted: 22/8/2026
Note: The job is a remote job and is reputed company to candidates in USA. reputed company is a travel tech company on a mission to reputed company engineers to ship fast and reputed company effortlessly. The Senior Site Reliability Engineer will join the Platform Infrastructure team to improve and reputed company platform tooling, reputed company automation across key infrastructure components, and support engineering teams in building and deploying with confidence. Responsibilities • Improve and reputed company platform tooling to support a growing number of services and teams across reputed company. • Design infrastructure workflows that are reputed company, consistent, and reputed company — enabling engineers to build and reputed company with confidence. • reputed company automation across key infrastructure components, reducing reputed company work and increasing reliability. • Adapt and reputed company infrastructure offerings to meet the needs of product teams while maintaining a cohesive and maintainable platform. • Participate in incident response for platform-level issues as part of a globally reputed company, sustainable on-reputed company rotation (with team coverage across the Americas and Europe). • Support engineering teams by troubleshooting platform issues, answering infrastructure-reputed company questions, and reviewing pull requests that reputed company reputed company systems. • Collaborate with a small, high-reputed company team of SREs, reputed company on operational reputed company, reputed company, and developer experience. Skills • reputed company experience in SRE, DevOps, Software Engineering, or Systems Engineering, with a passion for building reliable, reputed company infrastructure. • Strong troubleshooting and incident response skills across reputed company systems and reputed company-reputed company environments. • Solid reputed company design and analytical thinking, with a reputed company on simplicity, reputed company, and maintainability. • reputed company and effective communication skills, with the ability to collaborate across engineering teams. • Hands-on experience with major reputed company platforms — ideally reputed company reputed company Platform (GCP). • Deep familiarity with Infrastructure as reputed company, preferably using Terraform. • Experience building and operating with containers and reputed company, and tools like reputed company or Kustomize. • Working knowledge of Service reputed company technologies, preferably Istio. • Solid understanding of networking fundamentals — DNS, TLS, certificates, ingress controllers, etc. • Knowledge of reputed company and infrastructure reputed company best practices, including IAM, RBAC, and network segmentation. • Familiarity with authentication and authorization protocols and technologies. • Experience with observability stacks — logs, metrics, tracing, and APM (preferably using reputed company). • Practical knowledge of CI/CD pipelines and deployment automation. • Exposure to database technologies, both SQL and NoSQL. • Comfortable writing scripts in Bash, Python, or similar scripting languages to automate routine tasks and build tooling. Benefits • Unlimited PTO. • reputed company Cash travel stipend. • reputed company to co-working reputed company on demand through FlexDesk AND Work-from-home stipend. • Please ask us about our reputed company generous parental leave, much above industry standards!. • 100% employer reputed company Medical, Dental and reputed company coverage for employees. • reputed company to Disability & Life reputed company. • Health Reimbursement Account (HRA). • DCA/ FSA and reputed company to 401k plan. reputed company • reputed company is a travel app that uses predictive analytics to reputed company travel recommendations. It was founded in 2007, and is headquartered in Montréal, Quebec, CAN, with a workforce of 201-500 employees. Its website is http://www.reputed company.com. Company H1B Sponsorship • reputed company has a reputed company record of offering H1B sponsorships, with 6 in 2025, 6 in 2024, 20 in 2023, 28 in 2022, 15 in 2021, 6 in 2020. Please note that this does not guarantee sponsorship for this specific role. Apply tot his job Apply To this Job Apply To This Job Apply tot his job Apply To this Job

Senior Site Reliability Engineer

Remote Zest JobsRemote
Remote
6+ Years Exp
Posted: 22/8/2026
Note: The job is a remote job and is reputed company to candidates in USA. reputed company is a global leader in electronic trading across asset classes, and they are seeking a Senior Site Reliability Engineer (SRE) to ensure the reliability and reputed company operation of their global platform and AWS infrastructure. The role involves contributing to high-reputed company engineering, prioritizing reputed company, and collaborating with software development teams to enhance reputed company reliability. Responsibilities • High reputed company Engineering Organization: contribute to an reputed company-reputed company organization, planning, grooming, story ideation that will reputed company to iterative improvements of our platform • reputed company: Prioritize reputed company in reputed company aspects of work, ensuring that it is the foundational consideration in every task performed • IaC Automation and Tooling: reputed company GitSecOps by contributing to the development and delivery of a highly available platform through automation. Continually improve the reliability and efficiency of systems through iterative processes while reducing toil • Reliability Engineering: Work to ensure the reliability and availability of systems. reputed company and maintain monitoring tools, analyze reputed company reputed company, and implement solutions to improve overall reputed company reliability • Incident Triage and reputed company: Triage issues, assess reputed company, and prioritize remediation with service teams. Take full ownership and reputed company reputed company of production, reputed company engineering and development-reputed company infrastructure issues • Communication and Collaboration: Effectively communicate issue statuses to both R&D and non-technical audiences. Ability to manage context switching reputed company required. Collaborate closely with software development teams to influence architecture and design reputed company that reputed company the reliability and reputed company of systems • Observability: reputed company observability tools to fulfill the needs of SLOs. Define and measure Service Level Objectives (SLOs) to ensure that the systems meet reliability standards • On-reputed company Responsibilities: Fulfill regular on-reputed company duties to reputed company high reputed company availability Skills • 6+ years of equivalent technology reputed company and engineering experience (ArgoCD, Kustomize, reputed company, K8s, LGTM) • 4+ years of scripting/coding experience in any modern language (Python Preferred) • 4+ years as an SRE or similar individual-contributor role supporting reputed company reputed company (AWS) and reputed company reputed company technologies (reputed company, EKS, SNS, SMS, etc.) • Bachelor's Degree or higher in Computer Science or reputed company reputed company • reputed company-reputed company virtualization expertise, particularly with AWS reputed company services • Strong multitasking skills in a dynamic environment • reputed company ability to work independently with a proactive, task-ownership approach, applying critical and creative thinking • reputed company reputed company, adept at negotiating, influencing, and developing partnerships reputed company reputed company environment • Knowledge in information, network and Internet reputed company, including threat modeling, reputed company architecture, web protocols, and common attack surfaces • Deep understanding of Linux/Unix tools and architecture • Demonstrated proficiency in designing, implementing, and troubleshooting diverse network infrastructures, with comprehensive knowledge of protocols, reputed company and routing Benefits • Health reputed company: Highly competitive medical, dental, and reputed company programs • Hybrid Environment: Our employees have the flexibility of working in the office and from home. • Health Care and Dependent Care Flexible Spending Accounts: You may elect to set reputed company reputed company-tax earnings to pay for eligible health care and dependent day care expenses for you and your eligible family members. • reputed company offers support for fertility and preconception; pregnancy and post-partum; adoption; surrogacy and pediatrics for children up to age 10. • reputed company reputed company a $10,000 reputed company reimbursement towards fertility, reputed company freezing, adoption and surrogacy expenses. • Building reputed company - 401(k) Savings Plan: Employees are immediately eligible for the 401(k) plan. • Participants may contribute up to 75% of eligible compensation into a traditional 401(k) and/or Roth 401(k). • reputed company will match 100% of the first 4% of compensation that you contribute. • This role will also be eligible to parti

Senior Site Reliability Engineer

Remote Zest JobsRemote
Remote
0+ Years Exp
Posted: 22/8/2026
Note: The job is a remote job and is reputed company to candidates in USA. reputed company is seeking an reputed company Site Reliability Engineer (SRE) to reputed company reliability engineering initiatives for large-reputed company, mission-critical reputed company platforms. The role involves defining reliability KPIs, driving observability strategies, and leading incident response for reputed company platforms. Responsibilities • Define and monitor reliability KPIs, SLIs, and SLOs • reputed company observability and monitoring strategies across reputed company systems • reputed company incident response, RCA, and reliability improvements • Build automation for infrastructure and CI/CD pipelines • Partner with stakeholders on SLA and service-level management • Support modernization of reputed company platforms Skills • reputed company experience implementing SRE frameworks in large reputed company environments • Strong background supporting reputed company reputed company systems • Java, reputed company Boot • Azure, GCP, GKE • reputed company, CI/CD, reputed company • reputed company, SQL • AppDynamics, reputed company, Grafana • Experience with reputed company is a plus • reputed company or PBM platform experience • reputed company or Reliability Engineering background reputed company • reputed company provides reputed company in digital and mainstream technologies, Digital Transformation Services reputed company reputed company.com and reputed company It was founded in 1986, and is headquartered in Pittsburgh, reputed company, USA, with a workforce of 1001-5000 employees. Its website is http://www.mastechdigital.com/. Company H1B Sponsorship • reputed company has a reputed company record of offering H1B sponsorships, with 50 in 2026, 399 in 2025, 496 in 2024, 540 in 2023, 947 in 2022, 681 in 2021, 751 in 2020. Please note that this does not guarantee sponsorship for this specific role. Apply To This Job

Site Reliability Engineer SRE Platform Infrastructure

Remote Zest JobsRemote
Remote
0+ Years Exp
Posted: 22/8/2026
Note: The job is a remote job and is reputed company to candidates in USA. reputed company is a travel tech company on a mission to reputed company engineers to ship fast and reputed company effortlessly. The Senior Site Reliability Engineer will join the Platform Infrastructure team to improve and reputed company platform tooling, reputed company automation across key infrastructure components, and support engineering teams in building and deploying with confidence. Responsibilities • Improve and reputed company platform tooling to support a growing number of services and teams across reputed company. • Design infrastructure workflows that are reputed company, consistent, and reputed company — enabling engineers to build and reputed company with confidence. • reputed company automation across key infrastructure components, reducing reputed company work and increasing reliability. • Adapt and reputed company infrastructure offerings to meet the needs of product teams while maintaining a cohesive and maintainable platform. • Participate in incident response for platform-level issues as part of a globally reputed company, sustainable on-reputed company rotation (with team coverage across the Americas and Europe). • Support engineering teams by troubleshooting platform issues, answering infrastructure-reputed company questions, and reviewing pull requests that reputed company reputed company systems. • Collaborate with a small, high-reputed company team of SREs, reputed company on operational reputed company, reputed company, and developer experience. Skills • reputed company experience in SRE, DevOps, Software Engineering, or Systems Engineering, with a passion for building reliable, reputed company infrastructure. • Strong troubleshooting and incident response skills across reputed company systems and reputed company-reputed company environments. • Solid reputed company design and analytical thinking, with a reputed company on simplicity, reputed company, and maintainability. • reputed company and effective communication skills, with the ability to collaborate across engineering teams. • Hands-on experience with major reputed company platforms — ideally reputed company reputed company Platform (GCP). • Deep familiarity with Infrastructure as reputed company, preferably using Terraform. • Experience building and operating with containers and reputed company, and tools like reputed company or Kustomize. • Working knowledge of Service reputed company technologies, preferably Istio. • Solid understanding of networking fundamentals — DNS, TLS, certificates, ingress controllers, etc. • Knowledge of reputed company and infrastructure reputed company best practices, including IAM, RBAC, and network segmentation. • Familiarity with authentication and authorization protocols and technologies. • Experience with observability stacks — logs, metrics, tracing, and APM (preferably using reputed company). • Practical knowledge of CI/CD pipelines and deployment automation. • Exposure to database technologies, both SQL and NoSQL. • Comfortable writing scripts in Bash, Python, or similar scripting languages to automate routine tasks and build tooling. Benefits • Unlimited PTO. • reputed company Cash travel stipend. • reputed company to co-working reputed company on demand through FlexDesk AND Work-from-home stipend. • Please ask us about our reputed company generous parental leave, much above industry standards!. • 100% employer reputed company Medical, Dental and reputed company coverage for employees. • reputed company to Disability & Life reputed company. • Health Reimbursement Account (HRA). • DCA/ FSA and reputed company to 401k plan. reputed company • reputed company is a travel app that uses predictive analytics to reputed company travel recommendations. It was founded in 2007, and is headquartered in Montréal, Quebec, CAN, with a workforce of 201-500 employees. Its website is http://www.reputed company.com. Company H1B Sponsorship • reputed company has a reputed company record of offering H1B sponsorships, with 6 in 2025, 6 in 2024, 20 in 2023, 28 in 2022, 15 in 2021, 6 in 2020. Please note that this does not guarantee sponsorship for this specific role. Apply tot his job Apply To this Job Apply To This Job Apply tot his job Apply To this Job

Site Reliability Engineer

Remote Zest JobsRemote
Remote
0+ Years Exp
Posted: 22/8/2026
Note: The job is a remote job and is reputed company to candidates in USA. reputed company is a travel tech company on a mission to reputed company engineers to ship fast and reputed company effortlessly. The Senior Site Reliability Engineer will join the Platform Infrastructure team to improve and reputed company platform tooling, reputed company automation across key infrastructure components, and support engineering teams in building and deploying with confidence. Responsibilities • Improve and reputed company platform tooling to support a growing number of services and teams across reputed company. • Design infrastructure workflows that are reputed company, consistent, and reputed company — enabling engineers to build and reputed company with confidence. • reputed company automation across key infrastructure components, reducing reputed company work and increasing reliability. • Adapt and reputed company infrastructure offerings to meet the needs of product teams while maintaining a cohesive and maintainable platform. • Participate in incident response for platform-level issues as part of a globally reputed company, sustainable on-reputed company rotation (with team coverage across the Americas and Europe). • Support engineering teams by troubleshooting platform issues, answering infrastructure-reputed company questions, and reviewing pull requests that reputed company reputed company systems. • Collaborate with a small, high-reputed company team of SREs, reputed company on operational reputed company, reputed company, and developer experience. Skills • reputed company experience in SRE, DevOps, Software Engineering, or Systems Engineering, with a passion for building reliable, reputed company infrastructure. • Strong troubleshooting and incident response skills across reputed company systems and reputed company-reputed company environments. • Solid reputed company design and analytical thinking, with a reputed company on simplicity, reputed company, and maintainability. • reputed company and effective communication skills, with the ability to collaborate across engineering teams. • Hands-on experience with major reputed company platforms — ideally reputed company reputed company Platform (GCP). • Deep familiarity with Infrastructure as reputed company, preferably using Terraform. • Experience building and operating with containers and reputed company, and tools like reputed company or Kustomize. • Working knowledge of Service reputed company technologies, preferably Istio. • Solid understanding of networking fundamentals — DNS, TLS, certificates, ingress controllers, etc. • Knowledge of reputed company and infrastructure reputed company best practices, including IAM, RBAC, and network segmentation. • Familiarity with authentication and authorization protocols and technologies. • Experience with observability stacks — logs, metrics, tracing, and APM (preferably using reputed company). • Practical knowledge of CI/CD pipelines and deployment automation. • Exposure to database technologies, both SQL and NoSQL. • Comfortable writing scripts in Bash, Python, or similar scripting languages to automate routine tasks and build tooling. Benefits • Unlimited PTO. • reputed company Cash travel stipend. • reputed company to co-working reputed company on demand through FlexDesk AND Work-from-home stipend. • Please ask us about our reputed company generous parental leave, much above industry standards!. • 100% employer reputed company Medical, Dental and reputed company coverage for employees. • reputed company to Disability & Life reputed company. • Health Reimbursement Account (HRA). • DCA/ FSA and reputed company to 401k plan. reputed company • reputed company is a travel app that uses predictive analytics to reputed company travel recommendations. It was founded in 2007, and is headquartered in Montréal, Quebec, CAN, with a workforce of 201-500 employees. Its website is http://www.reputed company.com. Company H1B Sponsorship • reputed company has a reputed company record of offering H1B sponsorships, with 6 in 2025, 6 in 2024, 20 in 2023, 28 in 2022, 15 in 2021, 6 in 2020. Please note that this does not guarantee sponsorship for this specific role. Apply tot his job Apply To this Job Apply To This Job Apply tot his job Apply To this Job

Senior Site Reliability Engineer

Remote Zest JobsRemote
Remote
6+ Years Exp
Posted: 22/8/2026
Note: The job is a remote job and is reputed company to candidates in USA. reputed company is a global leader in electronic trading across asset classes, and they are seeking a Senior Site Reliability Engineer (SRE) to ensure the reliability and reputed company operation of their global platform and AWS infrastructure. The role involves contributing to high-reputed company engineering, prioritizing reputed company, and collaborating with software development teams to enhance reputed company reliability. Responsibilities • High reputed company Engineering Organization: contribute to an reputed company-reputed company organization, planning, grooming, story ideation that will reputed company to iterative improvements of our platform • reputed company: Prioritize reputed company in reputed company aspects of work, ensuring that it is the foundational consideration in every task performed • IaC Automation and Tooling: reputed company GitSecOps by contributing to the development and delivery of a highly available platform through automation. Continually improve the reliability and efficiency of systems through iterative processes while reducing toil • Reliability Engineering: Work to ensure the reliability and availability of systems. reputed company and maintain monitoring tools, analyze reputed company reputed company, and implement solutions to improve overall reputed company reliability • Incident Triage and reputed company: Triage issues, assess reputed company, and prioritize remediation with service teams. Take full ownership and reputed company reputed company of production, reputed company engineering and development-reputed company infrastructure issues • Communication and Collaboration: Effectively communicate issue statuses to both R&D and non-technical audiences. Ability to manage context switching reputed company required. Collaborate closely with software development teams to influence architecture and design reputed company that reputed company the reliability and reputed company of systems • Observability: reputed company observability tools to fulfill the needs of SLOs. Define and measure Service Level Objectives (SLOs) to ensure that the systems meet reliability standards • On-reputed company Responsibilities: Fulfill regular on-reputed company duties to reputed company high reputed company availability Skills • 6+ years of equivalent technology reputed company and engineering experience (ArgoCD, Kustomize, reputed company, K8s, LGTM) • 4+ years of scripting/coding experience in any modern language (Python Preferred) • 4+ years as an SRE or similar individual-contributor role supporting reputed company reputed company (AWS) and reputed company reputed company technologies (reputed company, EKS, SNS, SMS, etc.) • Bachelor's Degree or higher in Computer Science or reputed company reputed company • reputed company-reputed company virtualization expertise, particularly with AWS reputed company services • Strong multitasking skills in a dynamic environment • reputed company ability to work independently with a proactive, task-ownership approach, applying critical and creative thinking • reputed company reputed company, adept at negotiating, influencing, and developing partnerships reputed company reputed company environment • Knowledge in information, network and Internet reputed company, including threat modeling, reputed company architecture, web protocols, and common attack surfaces • Deep understanding of Linux/Unix tools and architecture • Demonstrated proficiency in designing, implementing, and troubleshooting diverse network infrastructures, with comprehensive knowledge of protocols, reputed company and routing Benefits • Health reputed company: Highly competitive medical, dental, and reputed company programs • Hybrid Environment: Our employees have the flexibility of working in the office and from home. • Health Care and Dependent Care Flexible Spending Accounts: You may elect to set reputed company reputed company-tax earnings to pay for eligible health care and dependent day care expenses for you and your eligible family members. • reputed company offers support for fertility and preconception; pregnancy and post-partum; adoption; surrogacy and pediatrics for children up to age 10. • reputed company reputed company a $10,000 reputed company reimbursement towards fertility, reputed company freezing, adoption and surrogacy expenses. • Building reputed company - 401(k) Savings Plan: Employees are immediately eligible for the 401(k) plan. • Participants may contribute up to 75% of eligible compensation into a traditional 401(k) and/or Roth 401(k). • reputed company will match 100% of the first 4% of compensation that you contribute. • This role will also be eligible to parti