Senior Application and Operation Support Engineer
OPTIVEUM sp. z o.o. Łódź, Łódź Voivodeship, Poland · PLN 218K/yr
Outsourcing and Offshoring Consulting · 2-10 employees
About the role
The role involves managing production incidents, performing root cause analysis, and ensuring the stability and performance of production environments. You will also oversee CI/CD deployment pipelines and collaborate with development teams to improve operational observability and automation.
What they look for
Requirements
Candidates must have at least 5 years of experience in IT operations or application support and 2 years of experience within an ITIL framework. Proficiency in log analysis tools, Kubernetes, Linux, and relational databases is essential for this position.
Full description
Senior Application and Operation Support Engineer
Location: Warsaw, Gdynia, or Łódź (Hybrid mode: 25% remote, 75% on-site)
Rate: Up to 105 PLN / hour (B2B)
Contracting Party: Optiveum (B2B cooperation agreement signed directly with us)
About the role:
We are currently looking for a Senior Application and Operation Support Engineer to join the IT operations team of one of our trusted clients. In this role, you will be the guardian of production environments, ensuring their stability, performance, and continuous improvement.
Please note that Optiveum is the recruitment and contracting party for this position, meaning you will sign the cooperation agreement directly with us while working on our partner's project.
Your day-to-day focus:
- Incident & Problem Management: Own the Root Cause Analysis (RCA) process for production incidents—diagnose, resolve, and put preventive measures in place so issues don't recur.
- Production Monitoring & Support: Continuously monitor service health, detect anomalies early, and act before they become full-blown incidents.
- Deployment Execution: Trigger and oversee release deployments through existing CI/CD pipelines; troubleshoot failed deployments and coordinate rollbacks when needed.
- Environment Oversight: Keep Pre-Production and Production environments stable and aligned, ensuring they behave as expected day to day.
- Operational Improvement: Identify recurring pain points and propose automation or tooling to reduce toil. Improve observability coverage (dashboards, alerts, log queries) and contribute to disaster recovery drills.
- Knowledge Management: Document operational procedures, known issues, and resolution steps to build a reliable knowledge base.
- Cross-team Collaboration: Work shoulder-to-shoulder with development and platform teams to triage issues, clarify operational requirements, and close the feedback loop between prod and dev.
Must-have knowledge and experience:
- 5+ years in IT operations, application support (L2/L3), or a similar production-facing role.
- Proven track record of owning incidents end-to-end—from alert to RCA to prevention.
- 2+ years working within an ITIL framework (incident, problem, and change management).
- Experience working in Agile delivery environments alongside development teams.
- Advanced English communication skills—ability to explain technical issues clearly to both engineers and non-technical stakeholders.
Must-have technical skills:
- Proficiency with log analysis and alerting tools (Splunk, Apica, Sysdig).
- Experience with Observability tooling (Prometheus, Grafana—reading dashboards, tuning alerts).
- Comfortable operating services running on Kubernetes and Linux CLI (checking pod health, reading logs, triggering restarts).
- Familiarity with Jenkins pipelines for executing and troubleshooting deployments, and proficiency in GIT.
- Experience with relational databases (Oracle, DB2)—querying, interpreting execution plans, and identifying data-related incidents.
- Working knowledge of Spring/Hibernate application behavior, Kafka message flows, and XML/JSON payloads to effectively trace issues through the stack.
Nice-to-have:
- Experience with Helm deployments.
- Java/J2EE development background (highly beneficial for reading stack traces and collaborating with devs).
- IBM DataStage operational experience.
- Scripting skills (Bash, Python) for automating repetitive operational tasks.
- Ansible knowledge for applying configuration changes in controlled scenarios.
Sound like a match? Apply today and join the Optiveum network!
Similar roles
-
Technical Support Engineer – Dutch Speaker
Dell Technologies Morocco
-
Junior Technical Support Engineer
Miratech Bengaluru, Karnataka, India
-
Application Support Engineer
Inetum Lisbon, Portugal
-
IT Support Engineer (M/F/D)
IFS Düsseldorf, North Rhine-Westphalia, Germany
-
Technical Support Engineer 1 (Commerce IT)
Choctaw Nation of Oklahoma Durant, Oklahoma, United States
-
Sales Support Engineer
Emerson Székesfehérvár, Central Transdanubia, Hungary