via Teamtailor
$60K - 90K a year
Manage production incidents, perform root cause analysis, ensure production stability, oversee CI/CD deployments, maintain documentation, and collaborate with development teams.
5+ years IT operations or application support experience, 2+ years ITIL, proficiency with monitoring tools, Kubernetes, Linux, databases, and advanced English communication.
Senior Application and Operation Support Engineer Location: Łódź, Gdańsk, Warsaw, or Tricity Work Model: Hybrid (25% remote) Rate: Up to 110 PLN/h (B2B) Contracting Party: Optiveum About the Role We are Optiveum, and we are currently looking for an experienced Senior Application and Operation Support Engineer to join our client's team. In this role, you will be crucial in ensuring the stability, performance, and continuous improvement of production and pre-production environments. You will sign a B2B cooperation agreement directly with Optiveum. Your Responsibilities Incident & Problem Management: Own the RCA process for production incidents—diagnose, resolve, and implement preventive measures. Production Monitoring & Support: Continuously monitor service health, detect anomalies early, and proactively prevent incidents. Deployment Execution: Trigger and oversee release deployments through existing CI/CD pipelines; troubleshoot failures and coordinate rollbacks when necessary. Environment Oversight: Maintain the stability and alignment of Pre-Production and Production environments. Runbook & Knowledge Management: Document operational procedures, known issues, and resolution steps to build a reliable team knowledge base. Operational Improvement: Identify recurring pain points and propose automation to reduce manual toil. Improve observability coverage (dashboards, alerts, log queries) and contribute to disaster recovery drills. Cross-team Collaboration: Work closely with development and platform teams to triage issues, clarify requirements, and close the feedback loop between production and development. Must-Have Requirements 5+ years of experience in IT operations, application support (L2/L3), or a similar production-facing role. Proven track record of owning incidents end-to-end (from alert to RCA to prevention). 2+ years of experience working within an ITIL framework (incident, problem, change management). Experience working in Agile delivery environments alongside development teams. Advanced English communication skills, with the ability to explain technical issues clearly to both engineers and non-technical stakeholders. Technical Skills Required Monitoring & Observability: Proficiency with Splunk, Apica, Sysdig, Prometheus, and Grafana (reading dashboards, tuning alerts). Infrastructure: Comfort operating services on Kubernetes and Linux CLI (e.g., checking pod health, reading logs, triggering restarts). CI/CD & Version Control: Familiarity with Jenkins pipelines for deployments and troubleshooting, and proficiency in GIT. Application & Data Layer: Experience with relational databases (Oracle, DB2)—querying, interpreting execution plans, and troubleshooting data incidents. Architecture Knowledge: Working knowledge of Spring/Hibernate behavior, Kafka message flows, and XML/JSON payloads to trace issues through the stack. Nice-to-Have Experience with Helm deployments. Java/J2EE development background (highly beneficial for reading stack traces and collaborating with devs). Operational experience with IBM DataStage. Scripting skills (Bash, Python) for automating repetitive tasks. Knowledge of Ansible for applying configuration changes.
This job posting was last updated on 8/24/2026