Important Company of the Sector•Chennai, Tamil Nadu, IN
0+ Years Exp
Posted: 24/8/2026
About Poshmark
Poshmark is a leading fashion resale marketplace powered by a vibrant, highly engaged community of buyers and sellers and real-time social experiences. Designed to make online selling fun, more social and easier than ever, Poshmark empowers its sellers to turn their closet into a thriving business and share their style with the world. Since its founding in 2011, Poshmark has grown its community to over 130 million users and generated over $10 billion in GMV, helping sellers realize billions in earnings, delighting buyers with deals and one-of-a-kind items, and building a more sustainable future for fashion. For more information, please visit www.poshmark.com, and for company news, visit newsroom.poshmark.com.
We are looking for a Site Reliability Engineer Intern to join our CloudOps/SRE team and support the reliability, availability, and scalability of Poshmark’s production systems. This internship is designed for candidates who are passionate about systems, automation, and cloud infrastructure and are eager to learn how large-scale web platforms are operated.
As an intern, you will work closely with experienced Site Reliability Engineers and Development teams to gain hands-on exposure to real-world production systems, monitoring, automation, and incident management processes.
6-Month Internship Accomplishments
- Gain a strong understanding of Poshmark’s technology stack, infrastructure, and core product functionality.
- Learn the fundamentals of Site Reliability Engineering, including monitoring, alerting, automation, and incident response.
- Get hands-on experience with cloud infrastructure and DevOps tools used within the CloudOps organization.
- Assist in building or enhancing automation scripts, dashboards, or internal tools.
- Contribute to small-scale projects under guidance from senior engineers.
- Observe and learn the on-call process and incident management workflows (no primary on-call ownership).
Responsibilities
- Assist in monitoring the health, performance, and availability of Poshmark’s services.
- Support senior engineers in deploying, maintaining, and improving infrastructure and services.
- Help create or update dashboards, alerts, and runbooks for production systems.
- Participate in troubleshooting basic production issues and post-incident reviews.
- Work closely with development teams to understand application behavior and operational needs.
- Follow best practices for reliability, security, and automation.
- Document learnings, processes, and tools clearly for team use.
Desired Skills & Qualifications
- Currently pursuing or recently completed a degree in Computer Science, Information Technology, or a related field.
- Basic understanding of Linux/UNIX operating systems.
- Familiarity with at least one programming or scripting language (Python, Ruby, Bash, or similar).
- Basic knowledge of cloud concepts (AWS/GCP/Azure fundamentals).
- Interest in DevOps, Site Reliability Engineering, or Cloud Infrastructure roles.
- Willingness to learn new tools, technologies, and operational practices.
- Good problem-solving skills and a proactive learning mindset.
- Ability to work effectively in a fast-paced, collaborative setting.
Technologies You’ll Be Exposed To
- Languages & Frameworks: Ruby, JavaScript, Node.js
- Web & Middleware: Nginx, Tomcat, HAProxy
- Datastores & Messaging: MongoDB, Redis, RabbitMQ, ElasticSearch
- Cloud: Amazon Web Services (EC2, RDS, S3, CloudFront, etc.)
- DevOps & SRE Tools: Terraform, Jenkins, Datadog, Kubernetes, Docker, Ansible
Lead Site Reliability Engineer SRE
FIS•Chennai, Tamil Nadu, IN
12+ Years Exp
Posted: 24/8/2026
Job Description:
Full time
Experienced (relevant combo of work and education)
Bachelor of Computer Engineering
Lead Site Reliability Engineer (SRE)
Are you curious, motivated, and forward-thinking? At FIS you’ll have the opportunity to work on some of the most challenging and relevant issues in financial services and technology. Our talented people empower us, and we believe in being part of a team that is open, collaborative, entrepreneurial, passionate and above all fun.
Work Location : Chennai (Two days in-office, Three days virtual)
Experience : 8 to 12 Years
What you will be doing:
• Defines meaningful SLIs, SLOs for owned services; tracks error budgets and burn rates.
• Advanced scripting, CI/CD, container orchestration (Docker/Kubernetes)
• Forecasts capacity for owned services; plans for growth
• Pattern Analysis and Correlation skillset
What you bring:
Must Have
• Strong Knowledge in Performance Monitoring Tools like Dynatrace, Splunk and ability to create Dashboards, Views and Alerts
• Knowledge in OS, Network, Middleware, Database, SSL, Load Balancer
• Must have knowledge in scripting language (Unix Shell, Windows Scripting, Python, Java, .NET…)
• Hands on with at least one of (Unix/Windows Scripting, Python, Java, C++, C#)
• Ability to Automate repetitive tasks - (Scripting, RPA, Power Automate, UiPath etc.)
• Ability to implement AI models in Problem Solving and reduce MTTR
• Experience with Virtualization / Containerization in a production environment (e.g. Kubernetes, etc.);
• Engineering Mindset with End to End view and provide solutions
• Passion for problem solving with strong analytical capabilities
Preferable to Have
• Banking / Payments Product Domain Knowledge
• Experience in Fintech working in Fintech Companies with rich Domain Experience
What we offer you:
• A work environment built on collaboration, flexibility and respect
• Competitive salary and attractive range of benefits designed to help support your lifestyle and wellbeing.
• Varied and challenging work to help you grow your technical skillset
FIS is committed to protecting the privacy and security of all personal information that we process in order to provide services to our clients. For specific information on how FIS protects personal information online, please see the Online Privacy Notice.
Sourcing Model
Recruitment at FIS works primarily on a direct sourcing model; a relatively small portion of our hiring is through recruitment agencies. FIS does not accept resumes from recruitment agencies which are not on the preferred supplier list and is not responsible for any related fees for resumes submitted to job postings, our employees, or any other part of our company.
#pridepass
Requirements:
Engineering Manager Site Reliability Engineering
Athenahealth•Chennai, Tamil Nadu, IN
10+ Years Exp
Posted: 24/8/2026
Grow your career internally or refer a friend to athenahealth!
Position Summary: We are seeking an Engineering Manager, Site Reliability Engineering (SRE), who is a hands-on technical people leader to lead the Service Operations Site Reliability Engineering team in Chennai within the Cloud Infrastructure Engineering (CIE) division. This role is responsible for driving reliability, observability, automation, and operational readiness across systems supporting Service Operations. The ideal candidate brings deep expertise in Linux infrastructure, observability platforms, infrastructure automation, incident management, and engineering leadership. This individual will partner closely with global engineering and operations teams to reduce toil, improve service reliability, and deliver scalable, resilient solutions that support athenahealth's mission of providing
About the Team: The Service Operations Site Reliability Engineering team is part of the Network Operations Center (NOC) organization and sits within the Cloud Infrastructure Engineering (CIE) division. The team is responsible for delivering highly available SaaS infrastructure, operational tooling, observability solutions, and automation capabilities that support Service Operations and Cloud Infrastructure teams. Working closely with R&D and Infrastructure stakeholders across India and the United States, the team focuses on improving operational excellence through automation, standardized onboarding, actionable monitoring, and continuous reduction of operational toil.
Essential Job Responsibilities:
• Lead, coach, mentor, and develop a team of Site Reliability and Infrastructure Engineers based in India.
• Remain technically hands-on by reviewing designs, guiding implementation efforts, troubleshooting complex issues, and contributing to technical solutions when required.
• Own team delivery across infrastructure management, observability, service onboarding, alerting, automation, and operational readiness initiatives.
• Drive observability strategy across metrics, logs, traces, synthetic monitoring, health checks, dashboards, and actionable alerting frameworks.
• Manage provisioning and lifecycle management of physical and virtual Linux systems using tools such as Puppet, Ansible, Terraform, and related automation platforms.
• Partner with engineering teams operating within SaaS, hybrid cloud, Kubernetes, and Amazon EKS environments to ensure complete monitoring and operational coverage.
• Identify, measure, and reduce operational toil through automation, self-service capabilities, documentation, and scalable operational processes.
• Lead Agile delivery practices including sprint planning, backlog prioritization, stakeholder communication, and continuous improvement activities.
Additional Job Responsibilities:
• Build and enhance monitoring integrations across platforms including New Relic, Prometheus, Alertmanager, OpenSearch, Grafana, Icinga, Unified Assurance, and related technologies.
• Establish Infrastructure-as-Code (IaC), Configuration-as-Code, Monitoring-as-Code, and Alerting-as-Code standards and practices.
• Improve alert quality by ensuring alerts contain actionable context, ownership, severity levels, routing information, and runbook references.
• Partner with NOC and Service Operations teams to standardize service onboarding, escalation management, operational handoffs, and response workflows.
• Manage hiring, onboarding, performance management, feedback, career development, and technical growth of direct reports.
• Participate in incident response activities, escalation reviews, post-incident analysis, and on-call planning processes.
• Develop and report operational metrics including alert quality, automation coverage, service health, onboarding throughput, toil reduction, and reliability improvements.
• Ensure operational excellence through comprehensive documentation, SOPs, runbooks, architecture diagrams, and support procedures while collaborating effectively with global teams.
Expected Education & Experience:
• Bachelor's degree in Computer Science, Information Technology, Engineering, or a related technical discipline; equivalent experience will also be considered.
• 10+ years of experience in Infrastructure Engineering, Site Reliability Engineering, Systems Engineering, Platform Engineering, or Technical Operations.
• 2+ years of experience managing or formally leading technical engineering teams.
• Strong hands-on experience administering, provisioning, and operating Linux systems in large-scale production environments.
• Proven experience with observability platforms, monitoring, logging, tracing, dashboarding, alerting, and synthetic monitoring solutions.
• Experience with Infrastructure-as-Code and configuration management tools such as Terraform, Puppet, Ansible, Chef, or similar technologies, along with scripting in Python, Go, Bash, Ruby, Java, or related languages.
• Experience supporting SaaS, hybrid cloud, Kubernetes/EKS environments, CI/CD pipelines, incident management, operational readiness, and modern engineering practices, with strong communication and stakeholder management skills.
Have you notified your current manager of your application?
Site Reliability Engineer
UPS•Chennai, Tamil Nadu, IN
4+ Years Exp
Posted: 23/8/2026
Before you apply to a job, select your language preference from the options available at the top right of this page.
Explore your next prospect at a Fortune Global 500 organization. Envision innovative possibilities, experience our rewarding culture, and work with talented teams that help you become better every day. We know what it takes to lead UPS into tomorrowpeople with a unique combination of skill + passion. If you have the qualities and drive to lead yourself or teams, there are roles ready to cultivate your skills and take you to the next level.
Job Description
Responsibilities:
System Reliability:
Ensure the reliability and uptime of critical services and infrastructure.
Google Cloud Expertise:
Design, implement, and manage cloud infrastructure using Google Cloud services.
Automation:
Develop and maintain automation scripts and tools to improve system efficiency and reduce manual intervention.
Monitoring and Incident Response:
Implement monitoring solutions and respond to incidents to minimize downtime and ensure quick recovery.
Collaboration:
Work closely with development and operations teams to improve system reliability and performance.
Capacity Planning:
Conduct capacity planning and performance tuning to ensure systems can handle future growth.
Documentation:
Create and maintain comprehensive documentation for system configurations, processes, and procedures.
Qualifications
Education: Bachelors degree in computer science, Engineering, or a related field.
Experience: 4+ years of experience in site reliability engineering or a similar role.
Skills
Proficiency in Google Cloud services (Compute Engine, Kubernetes Engine, Cloud Storage, BigQuery, Pub/Sub, etc.).
Familiarity with Google BI and AI/ML tools (Looker, BigQuery ML, Vertex AI, etc.)
Experience with automation tools (Terraform, Ansible, Puppet).
Familiarity with CI/CD pipelines and tools (Azure pipelines Jenkins, GitLab CI, etc.).
Strong scripting skills (Python, Bash, etc.).
K .
Senior App Developer Java Production Engineering
UPS•Chennai, Tamil Nadu, IN
8+ Years Exp
Posted: 23/8/2026
Before you apply to a job, select your language preference from the options available at the top right of this page.
Explore your next opportunity at a Fortune Global 500 organization. Envision innovative possibilities, experience our rewarding culture, and work with talented teams that help you become better every day. We know what it takes to lead UPS into tomorrow—people with a unique combination of skill + passion. If you have the qualities and drive to lead yourself or teams, there are roles ready to cultivate your skills and take you to the next level.
Job Description:
Role Summary
The Senior Software Development Engineer is responsible for ensuring production stability while contributing to engineering improvements, automation, and feature delivery. This role combines software engineering expertise with production ownership.
Responsibilities
• Own runtime health and operational stability of production services.
• Participate in on-call support rotations.
• Investigate production incidents and implement permanent fixes.
• Develop automation for monitoring, alerting, deployments, and incident remediation.
• Improve system reliability, scalability, and performance.
• Perform capacity planning and performance tuning.
• Support Couchbase and PostgreSQL databases.
• Leverage AI tools to improve engineering productivity and operational efficiency.
• Rotate between SRE and feature engineering teams based on business priorities.
Required Skills
• 8+ years of backend software engineering experience.
• Strong Java programming skills.
• Experience with REST APIs and microservices.
• Experience with messaging platforms (AMQ, MQ, Kafka).
• Azure or GCP cloud experience.
• Strong debugging and production troubleshooting skills.
• Experience with Couchbase and PostgreSQL.
• Knowledge of monitoring and logging platforms.
As part of the Site Reliability Engineering (SRE) organization, every team member plays a critical role in maintaining the reliability and availability of our production systems.
This role supports a 24x7 production environment supporting business-critical applications.
All engineers are expected to participate in an on-call support rotation, including evenings, weekends, and public holidays as scheduled.
During on-call periods, you will be expected to respond to production incidents within defined response time objectives, troubleshoot issues, and work collaboratively to restore service as quickly as possible.
Candidates should be comfortable diagnosing complex production issues across applications, cloud infrastructure, databases, messaging systems, and integrations.
Participation in post-incident reviews, root cause analysis (RCA), and implementation of permanent corrective actions is an essential part of the role.
Team members are expected to continuously improve operational processes through automation, monitoring enhancements, and proactive reliability improvements to reduce recurring incidents.
Flexibility to collaborate with globally distributed teams, including occasional overlap with US business hours, is expected to support production operations and major releases.
A strong sense of ownership, accountability, and customer focus is essential, as ensuring platform stability and business continuity is a shared responsibility across the entire team.
Employee Type:
Permanent
UPS is committed to providing a workplace free of discrimination, harassment, and retaliation.
Engineering Manager Site Reliability Engineering
Athenahealth•Chennai, Tamil Nadu, IN
10+ Years Exp
Posted: 23/8/2026
Grow your career internally or refer a friend to athenahealth!
Position Summary: We are seeking an Engineering Manager, Site Reliability Engineering (SRE), who is a hands-on technical people leader to lead the Service Operations Site Reliability Engineering team in Chennai within the Cloud Infrastructure Engineering (CIE) division. This role is responsible for driving reliability, observability, automation, and operational readiness across systems supporting Service Operations. The ideal candidate brings deep expertise in Linux infrastructure, observability platforms, infrastructure automation, incident management, and engineering leadership. This individual will partner closely with global engineering and operations teams to reduce toil, improve service reliability, and deliver scalable, resilient solutions that support athenahealth's mission of providing
About the Team: The Service Operations Site Reliability Engineering team is part of the Network Operations Center (NOC) organization and sits within the Cloud Infrastructure Engineering (CIE) division. The team is responsible for delivering highly available SaaS infrastructure, operational tooling, observability solutions, and automation capabilities that support Service Operations and Cloud Infrastructure teams. Working closely with R&D and Infrastructure stakeholders across India and the United States, the team focuses on improving operational excellence through automation, standardized onboarding, actionable monitoring, and continuous reduction of operational toil.
Essential Job Responsibilities:
• Lead, coach, mentor, and develop a team of Site Reliability and Infrastructure Engineers based in India.
• Remain technically hands-on by reviewing designs, guiding implementation efforts, troubleshooting complex issues, and contributing to technical solutions when required.
• Own team delivery across infrastructure management, observability, service onboarding, alerting, automation, and operational readiness initiatives.
• Drive observability strategy across metrics, logs, traces, synthetic monitoring, health checks, dashboards, and actionable alerting frameworks.
• Manage provisioning and lifecycle management of physical and virtual Linux systems using tools such as Puppet, Ansible, Terraform, and related automation platforms.
• Partner with engineering teams operating within SaaS, hybrid cloud, Kubernetes, and Amazon EKS environments to ensure complete monitoring and operational coverage.
• Identify, measure, and reduce operational toil through automation, self-service capabilities, documentation, and scalable operational processes.
• Lead Agile delivery practices including sprint planning, backlog prioritization, stakeholder communication, and continuous improvement activities.
Additional Job Responsibilities:
• Build and enhance monitoring integrations across platforms including New Relic, Prometheus, Alertmanager, OpenSearch, Grafana, Icinga, Unified Assurance, and related technologies.
• Establish Infrastructure-as-Code (IaC), Configuration-as-Code, Monitoring-as-Code, and Alerting-as-Code standards and practices.
• Improve alert quality by ensuring alerts contain actionable context, ownership, severity levels, routing information, and runbook references.
• Partner with NOC and Service Operations teams to standardize service onboarding, escalation management, operational handoffs, and response workflows.
• Manage hiring, onboarding, performance management, feedback, career development, and technical growth of direct reports.
• Participate in incident response activities, escalation reviews, post-incident analysis, and on-call planning processes.
• Develop and report operational metrics including alert quality, automation coverage, service health, onboarding throughput, toil reduction, and reliability improvements.
• Ensure operational excellence through comprehensive documentation, SOPs, runbooks, architecture diagrams, and support procedures while collaborating effectively with global teams.
Expected Education & Experience:
• Bachelor's degree in Computer Science, Information Technology, Engineering, or a related technical discipline; equivalent experience will also be considered.
• 10+ years of experience in Infrastructure Engineering, Site Reliability Engineering, Systems Engineering, Platform Engineering, or Technical Operations.
• 2+ years of experience managing or formally leading technical engineering teams.
• Strong hands-on experience administering, provisioning, and operating Linux systems in large-scale production environments.
• Proven experience with observability platforms, monitoring, logging, tracing, dashboarding, alerting, and synthetic monitoring solutions.
• Experience with Infrastructure-as-Code and configuration management tools such as Terraform, Puppet, Ansible, Chef, or similar technologies, along with scripting in Python, Go, Bash, Ruby, Java, or related languages.
• Experience supporting SaaS, hybrid cloud, Kubernetes/EKS environments, CI/CD pipelines, incident management, operational readiness, and modern engineering practices, with strong communication and stakeholder management skills.
Have you notified your current manager of your application?
Associate Project Engineer Cyber Security
Hitachi Careers•Chennai, Tamil Nadu, IN
0+ Years Exp
Posted: 23/8/2026
Job Description:
The opportunity:
The technical marketing engineer for Mission Critical telecommunication Solutions (MCS) has the global responsibility to enable the Pre-Sales & Sales community of the different regional HUBs to understand technical market requirements for wired telecommunication networks and ensure customer interaction in line with global solution/product strategy. Support sales organizations in driving sales by your technical expertise. Provide relevant customer & market inputs to product management and R&D activities, ensuring market alignment and relevance.
How you'll make an impact:
• You will support Global Projects by adapting HVDC Cyber security for SCADA & HMI, develop architecture & functional descriptions for Functions / Solutions for future HVDC in cyber security technologies.
• You will prepare & perform test case scenarios & participate in FST, support projects in resolving the issues related to Cyber security Functions. You will coordinate with different stakeholders across the business units to get inputs to optimise the Cyber security solutions in HVDC.
• You will design & develop a secured network architecture for SCADA system with advanced cyber security features. You will evaluate and strengthen the security of any connections to the SCADA network.
• You will monitor and validate third party security patches to ensure that reliability of the system is maintained, implement the security features provided by device and system vendors.
• You will establish strong controls over any medium that is used as a backdoor into the SCADA network, implement internal and external intrusion detection systems in the SCADA network.
• You will perform technical audits of SCADA devices and networks, and any other connected networks, to identify security concerns.
• You will conduct physical security surveys and assess all remote sites connected to the SCADA network to evaluate their security, maintain the test environment with updated software and hardware for current and future use.
• You will backup/ Image handling on servers and workstations and patch management.
• You will be responsible to ensure compliance with applicable external and internal regulations, procedures, and guidelines.
• Living Hitachi Energy's core values of safety and integrity, which means taking responsibility for your own actions while caring for your colleagues and the business.
Your background:
• You hold a bachelor’s degree in Computer science/IT/ECE/EEE with a minimum knowledge in networking & protections.
• You must have knowledge & skills in Networking, security, Anti-Virus.
• Basic knowledge on IEEE / IEC standards.
• Basic Knowledge on C/C++ and experience in SCADA projects will be an added advantage.
• Knowledge & Experience MS Office: Word, Excel.
• Self-starter caliber who could own tasks through to completion.
• Strong attention to detail.
• Excellent written and verbal communication skills.
Accessibility and reasonable accommodation
Qualified individuals with a disability may request a reasonable accommodation if you are unable or limited in your ability to use or access the Hitachi Energy career site as a result of your disability. You may request reasonable accommodations by completing a general inquiry form on our website. Please include your contact information and specific details about your required accommodation to support you during the job application process.
This is solely for job seekers with disabilities requiring accessibility assistance or an accommodation in the job application process. Messages left for other purposes will not receive a response.
Use of Al and automated tools in recruitment
As part of our recruitment process, Hitachi Energy uses digital and automated tools, including Al-supported solutions, to assist with activities such as application screening, job matching, and interview scheduling. These tools are designed to support our recruiters and do not replace human decision-making. Candidate data is processed in accordance with applicable data protection and employment laws as well as Hitachi's Global Data Privacy Notice.
Background Screening and Security Checks
As part of the hiring process, Hitachi Energy conducts pre-employment background checks that may include verification of employment history, education, criminal records, and other relevant information, in accordance with applicable laws.
For certain roles—particularly those involving access to sensitive information, financial responsibilities, client data, regulated environments, or security-sensitive functions—additional or more comprehensive background or security screenings may be required. These may include, but are not limited to, enhanced criminal history checks, credit history reviews (where legally permissible), sanctions screening, or other due diligence measures aligned with the responsibilities of the position.
The scope and depth of any background or security review will be determined based on the nature of the role and business necessity, and will always be conducted in compliance with applicable federal, state, and local laws. Candidates will be notified and, where required, asked to provide consent prior to the initiation of any such checks.
Site Reliability Engineer
UPS•Chennai, Tamil Nadu, IN
4+ Years Exp
Posted: 23/8/2026
Before you apply to a job, select your language preference from the options available at the top right of this page.
Explore your next prospect at a Fortune Global 500 organization. Envision innovative possibilities, experience our rewarding culture, and work with talented teams that help you become better every day. We know what it takes to lead UPS into tomorrowpeople with a unique combination of skill + passion. If you have the qualities and drive to lead yourself or teams, there are roles ready to cultivate your skills and take you to the next level.
Job Description
Responsibilities:
System Reliability:
Ensure the reliability and uptime of critical services and infrastructure.
Google Cloud Expertise:
Design, implement, and manage cloud infrastructure using Google Cloud services.
Automation:
Develop and maintain automation scripts and tools to improve system efficiency and reduce manual intervention.
Monitoring and Incident Response:
Implement monitoring solutions and respond to incidents to minimize downtime and ensure quick recovery.
Collaboration:
Work closely with development and operations teams to improve system reliability and performance.
Capacity Planning:
Conduct capacity planning and performance tuning to ensure systems can handle future growth.
Documentation:
Create and maintain comprehensive documentation for system configurations, processes, and procedures.
Qualifications
Education: Bachelors degree in computer science, Engineering, or a related field.
Experience: 4+ years of experience in site reliability engineering or a similar role.
Skills
Proficiency in Google Cloud services (Compute Engine, Kubernetes Engine, Cloud Storage, BigQuery, Pub/Sub, etc.).
Familiarity with Google BI and AI/ML tools (Looker, BigQuery ML, Vertex AI, etc.)
Experience with automation tools (Terraform, Ansible, Puppet).
Familiarity with CI/CD pipelines and tools (Azure pipelines Jenkins, GitLab CI, etc.).
Strong scripting skills (Python, Bash, etc.).
K .
Cloud DevOps Engineer
UPS•Chennai, Tamil Nadu, IN
5+ Years Exp
Posted: 20/8/2026
Before you apply to a job, select your language preference from the options available at the top right of this page.
Explore your next opportunity at a Fortune Global 500 organization. Envision innovative possibilities, experience our rewarding culture, and work with talented teams that help you become better every day. We know what it takes to lead UPS into tomorrow—people with a unique combination of skill + passion. If you have the qualities and drive to lead yourself or teams, there are roles ready to cultivate your skills and take you to the next level.
Job Description:
DevOps Platform & Cloud Operations Engineer Job Summary The DevOps Platform & Cloud Operations Engineer will partner with Enterprise Architecture, Enterprise Security, and Platform Operations teams to deliver secure, reliable, and standardized cloud platforms, CI/CD pipelines, release processes, and operational support. The role focuses on automation, governance, cloud operations, platform stability, and DevSecOps best practices. Key Responsibilities ● Build and maintain CI/CD pipelines using Azure DevOps, Jenkins, and automation frameworks. ● Automate build, deployment, testing, validation, and rollback processes. ● Implement and support Infrastructure as Code (IaC), environment standardization, and policy-as-code controls. ● Support Azure and GCP cloud environments, networking, identity management, and hybrid integrations. ● Manage release readiness, deployment coordination, post-release validation, and dependency management. ● Align platforms, pipelines, and deployment processes with Enterprise Architecture and Security standards. ● Implement CI/CD governance including branch protection, approvals, segregation of duties, secrets management, and compliance gates. ● Support vulnerability remediation, patching, certificate management, and access governance. ● Monitor platform health, investigate incidents, perform root cause analysis, and drive operational improvements. ● Develop automation scripts, reusable operational tools, and standard operating procedures to improve efficiency and scalability. ● Provide production support, environment troubleshooting, and operational continuity during critical incidents. Required Skills ● Azure DevOps, Jenkins, CI/CD Pipeline Engineering ● Infrastructure as Code (Terraform preferred) ● Azure & GCP Cloud Platforms ● Cloud Networking, Identity & Access Management ● Release Management & Deployment Automation ● DevSecOps, Security Controls & Compliance ● Monitoring, Logging & Incident Management ● PowerShell, Bash, Python, or Automation Scripting ● Git, Version Control, and Agile Delivery Practices Preferred Qualifications ● 5+ years in DevOps, Cloud Engineering, Platform Operations, or Site Reliability Engineering. ● Azure and/or GCP certifications preferred. ● Experience working with Enterprise Architecture, Security, and Governance teams. ● Strong troubleshooting, operational support, and automation mindset.
Employee Type:
Permanent
UPS is committed to providing a workplace free of discrimination, harassment, and retaliation.
Cloud DevOps Engineer
UPS•Chennai, Tamil Nadu, IN
5+ Years Exp
Posted: 20/8/2026
Before you apply to a job, select your language preference from the options available at the top right of this page.
Explore your next opportunity at a Fortune Global 500 organization. Envision innovative possibilities, experience our rewarding culture, and work with talented teams that help you become better every day. We know what it takes to lead UPS into tomorrow—people with a unique combination of skill + passion. If you have the qualities and drive to lead yourself or teams, there are roles ready to cultivate your skills and take you to the next level.
Job Description:
DevOps Platform & Cloud Operations Engineer Job Summary The DevOps Platform & Cloud Operations Engineer will partner with Enterprise Architecture, Enterprise Security, and Platform Operations teams to deliver secure, reliable, and standardized cloud platforms, CI/CD pipelines, release processes, and operational support. The role focuses on automation, governance, cloud operations, platform stability, and DevSecOps best practices. Key Responsibilities ● Build and maintain CI/CD pipelines using Azure DevOps, Jenkins, and automation frameworks. ● Automate build, deployment, testing, validation, and rollback processes. ● Implement and support Infrastructure as Code (IaC), environment standardization, and policy-as-code controls. ● Support Azure and GCP cloud environments, networking, identity management, and hybrid integrations. ● Manage release readiness, deployment coordination, post-release validation, and dependency management. ● Align platforms, pipelines, and deployment processes with Enterprise Architecture and Security standards. ● Implement CI/CD governance including branch protection, approvals, segregation of duties, secrets management, and compliance gates. ● Support vulnerability remediation, patching, certificate management, and access governance. ● Monitor platform health, investigate incidents, perform root cause analysis, and drive operational improvements. ● Develop automation scripts, reusable operational tools, and standard operating procedures to improve efficiency and scalability. ● Provide production support, environment troubleshooting, and operational continuity during critical incidents. Required Skills ● Azure DevOps, Jenkins, CI/CD Pipeline Engineering ● Infrastructure as Code (Terraform preferred) ● Azure & GCP Cloud Platforms ● Cloud Networking, Identity & Access Management ● Release Management & Deployment Automation ● DevSecOps, Security Controls & Compliance ● Monitoring, Logging & Incident Management ● PowerShell, Bash, Python, or Automation Scripting ● Git, Version Control, and Agile Delivery Practices Preferred Qualifications ● 5+ years in DevOps, Cloud Engineering, Platform Operations, or Site Reliability Engineering. ● Azure and/or GCP certifications preferred. ● Experience working with Enterprise Architecture, Security, and Governance teams. ● Strong troubleshooting, operational support, and automation mindset.
Employee Type:
Permanent
UPS is committed to providing a workplace free of discrimination, harassment, and retaliation.