Pitney Bowes

Senior Site Reliability Engineer

Pitney Bowes pune, Maharashtra, India

Software Development · 10,001+ employees

13 h ago
sre Senior (5-10 yrs) Full-time India
Create a free account to apply — email only, no card. You can also save this posting or score it against your profile with AI.

About the role

The Senior Site Reliability Engineer will manage and maintain mission-critical enterprise applications to ensure high availability and system performance. Responsibilities include implementing monitoring solutions, conducting root cause analysis, and automating infrastructure tasks to improve reliability.

What they look for

Site Reliability Engineering AWS DevOps Kubernetes Docker CI/CD Monitoring Incident Management Linux Windows Python Terraform Ansible Prometheus Grafana Splunk

Requirements

Candidates must have 4-6 years of relevant SRE experience with a strong background in AWS, DevOps tools, and container orchestration. A degree in computer science or a related field is required, along with excellent analytical and troubleshooting skills.

Benefits

Comprehensive benefits globally Wellbeing programs Inclusive environment

Full description

We’re hiring at Pitney Bowes, where top talent builds meaningful careers and lasting impact. We Move fast, Deliver excellence, and Win together…that’s The Pitney Bowes way. Here, how we work matters just as much as what we achieve.

We’re looking for people who:

  • Act with urgency, accountability, and purpose
  • Deliver high quality work with consistency and pride
  • Collaborate effectively and elevate those around them
  • Focus on outcomes that drive impact and growth

Job Description:

Join Pitney Bowes as a Senior Site Reliability Engineer

Years of experience: 4 – 6 years

Job Location – Pune

Impact

As a Senior Site Reliability Engineer, you will be part of our core SRE team that Supports and maintains mission critical enterprise applications. You will be involved in ensuring the availability of the Products and the infrastructure on which they are hosted.  You will have access to various monitoring and troubleshooting tools that you would be using to do an investigation and resolve issues. You would also be responsible for doing a Root Cause Analysis and take necessary Corrective Actions to prevent reoccurrence of issues. Being a Senior SRE you will be a person who brings fresh ideas, demonstrates a unique and informed viewpoint, and enjoys collaborating with a cross-functional team to develop real-world solutions and positive user experiences at every interaction.

The Job

  • Run the production environment with highest availability by effective monitoring and taking a holistic view of system health.
  • Build software and systems to manage platform infrastructure and applications.
  • Improve reliability, quality, and time-to-market of our suite of software solutions.
  • Measure and optimize system performance, with an eye toward pushing our capabilities forward, getting ahead of customer needs, and innovating to continually improve.
  • Provide primary support under Continuous Integration and Delivery (CI/CD) across SDLC phases for multiple large-distributed software applications and not limited to...
  • Monitoring
  • Alert configuration and development
  • Automations to reduce manual effort
  • Quick and reliable Incident Response
  • Infrastructure Provision, maintenance, and optimization (Cost Optimization for Advisory and above)
  • Deployment and Patching
  • Communicates effectively on risks, issues, and changes associated to the product with stakeholders and teams.
  • Work with team to identify and implement what/how aspects of effective monitors. Ensuring that application monitoring is done in the best possible way and alerts, monitors with maximum noise reduction.
  • Understanding and analysis of User Personas and find best ways to optimize and evolve automate monitoring.
  • Specifying Service Level Indicators and Objectives to stay above the committed SLA.
  • Responsible for identifying and creating key metrics such as QoS, Uptime, MTTR, MTBF while keeping an eye on performance. (Preparing and publishing for advisory and above)
  • Strong experience on various AWS services as well as DevOps tools
  • You will design and implement monitoring solutions using advanced tools like Sumologic, Dynatrace, Splunk, CloudWatch, Grafana, Prometheus, Nagios, PagerDuty, Opsgenie
  • Organize and facilitate outage management activities including outage status communication, problem detection and resolution.
  • Incident Management and Disaster Recovery
  • Documenting “Tribal” Knowledge and create/maintain the KB
  • Conducting Postmortem for Incidents and Root Cause Analysis
  • Partner with development and QA teams to improve services through rigorous testing and release procedures.
  • Participate in system design consulting, platform management, and capacity planning.
  • Providing On-Call Support and Issue Resolution
  • Demonstrate Ownership and accountability.

Qualifications & Skills required.

This is a critical service delivery role requiring experience with complex datacenter and cloud hosting environments. Pitney Bowes product hosting solutions leverage multiple technologies in complex data center and cloud environments that support multi-tiered high-availability applications. The role requires a talented self-directed and self-motivated individual with a strong work ethic and the following skills:

  • Graduate or Post-Graduate (preferably in computer science or related course)
  • 4-6 years relevant SRE experience with Product/Application Support and DevOps experience.
  • Excellent written and verbal communication, time management, and presentation skills.
  • Experience of SaaS based product/application support.
  • Good Experience in CI/CD tools like GIT and ARGO
  • Strong experience in Docker, Kubernetes, Orchestration and Microservices
  • Strong experience with AWS services and cloud technologies
  • Good exposure to Firewall concept and networking experience.
  • Good to have infrastructure as code experience (Ansible, Terraform, CloudFormation)
  • Experience in scripting languages (shell scripts, Perl, Python, PowerShell)
  • Experience with Prometheus, Grafana, Dynatrace, Splunk, SumoLogic, Nagios, Pagerduty and Opsgenie
  • Strong knowledge on Linux and Windows operating systems
  • Strong analytical and application troubleshooting/debugging skills
  • Experience with JIRA, Confluence, SharePoint
  • Working experience Agile Scrum methodology
  • Strong cross-functional collaboration skill is mandatory - with teams like Product Development, Product Management, Client Success, Field Services, IT and other operations groups including Sr. leadership team.
  • Strong sense of personal responsibility and accountability for delivering high quality work, both personally and at a team level 

About Pitney Bowes

Pitney Bowes (NYSE:PBI) is a global technology company providing commerce solutions that power billions of transactions. Clients around the world, including 90 percent of the Fortune 500, rely on the accuracy and precision delivered by Pitney Bowes solutions, analytics, and APIs in the areas of ecommerce fulfillment, shipping and returns; cross-border ecommerce; office mailing and shipping; presort services; and financing. For 100 years Pitney Bowes has been innovating and delivering technologies that remove the complexity of getting commerce transactions precisely right. For additional information visit Pitney Bowes at https://www.pitneybowes.com/in.

We will:

  • Provide the will: opportunity to grow and develop your career
  • Offer an inclusive environment that encourages diverse perspectives and ideas
  • Deliver challenging and unique opportunities to contribute to the success of a transforming organization
  • Offer comprehensive benefits globally (PB Benefits and Wellbeing Programs)

Pitney Bowes is an equal opportunity employer that values diversity and inclusiveness in the workplace.

All interested individuals must apply online.

Similar roles