About the role
Job Description
* Ensure the stability, availability, and performance of production applications. * Monitor, troubleshoot, and resolve incidents, problems, and performance issues. * Participate in major incident management (P1/P2), root cause analysis (RCA), and continuous service improvement initiatives. * Manage deployments, releases, and change requests following ITIL and DevOps best practices. * Collaborate with Development, Infrastructure, Scrum teams, and external providers to deliver sustainable solutions. * Implement and enhance monitoring and observability capabilities across production environments. * Maintain technical documentation and promote knowledge sharing within global support teams. * Participate in on-call rotations supporting business-critical applications.
Qualifications
* 6+ years of experience in Application Production Support, Infrastructure Operations, or similar environments. * Strong experience with Java Application Servers (Red Hat JBoss EAP), including performance optimization, heap dumps, and thread dump analysis. * Hands-on experience with OpenShift, Kubernetes, and Cloud environments. * Experience with API Gateway solutions (Axway, Apigee). * Strong knowledge of RHEL Linux. * Experience with monitoring and observability tools such as Dynatrace, Grafana, Prometheus, ELK, and Jaeger. * Experience with CI/CD and DevOps tools, including GitLab, Jenkins, ArgoCD, and Nexus. * Experience with Ansible and/or Terraform for automation and infrastructure management. * Working knowledge of SQL Server and PostgreSQL. * Experience working with ITIL, Agile, and DevOps methodologies. * Portuguese (Fluent) * English (Professional working proficiency)
Additional Information
* Availability to travel within Portugal when required; * Availability for occasional international travel; * Experience in multicultural and international environments is considered a plus.