Sr. SRE (Application Support + Dev-Ops + Automation)
Fulcrum Digital · Shankill, Leinster, Ireland
IT Services and IT Consulting · 1,001-5,000 employees
About the role
Manage day-to-day production support activities, including incident management, change requests, and root cause analysis. Build and maintain CI/CD pipelines, monitor application health, and execute performance testing to ensure system availability.
What they look for
Requirements
The role requires experience in managing ITIL-based service management processes and troubleshooting production environments. Proficiency in tools such as Jenkins, Splunk, Dynatrace, and Digital.ai Release is essential for this position.
Full description
Who are we
Fulcrum Digital is an agile and next-generation digital accelerating company providing digital transformation and technology services right from ideation to implementation. These services have applicability across a variety of industries, including banking & financial services, insurance, retail, higher education, food, healthcare, and manufacturing.
Requirements
- Manage day-to-day production support activities while ensuring high system availability and performance.
- Handle ITIL-based service management processes including:
- Incident Management
- Problem Management (PBI)
- Change Requests (CRQs)
- Service Requests
- Root Cause Analysis (RCA)
Key Responsibilities
- Build, maintain, and troubleshoot CI/CD pipelines using Jenkins.
- Manage application deployment and release orchestration using Digital.ai Release (XL Release/XLR).
- Monitor application health and infrastructure using Dynatrace and proactively resolve performance issues.
- Develop and maintain dashboards, alerts, and log analytics using Splunk.
- Execute performance and load testing using BlazeMeter and analyze test results to identify bottlenecks.
- Plan and execute SSL/TLS certificate renewals, ensuring uninterrupted service availability.
- Perform operating system, middleware, and application patching activities while minimizing downtime.
- Participate in disaster recovery planning, testing, and execution to ensure business continuity.
- Collaborate with development, infrastructure, security, and business teams for production releases and operational improvements.
- Prepare operational documentation, SOPs, runbooks, and post-implementation reviews.
- Participate in on-call support and production incident resolution.