Application Production Support Engineer

Posted Yesterday
Be an Early Applicant
Lisbon, PRT
Hybrid
Senior level
Information Technology • Consulting
The Role
Manage and support production business applications to ensure availability, stability, and performance. Lead incident and problem management, RCAs, deployments, and on-call rotations. Implement and optimize monitoring/observability, collaborate with DevOps and development teams, automate operations, and maintain runbooks and documentation to improve platform reliability and SLAs.
Summary Generated by Built In
Company Description

Inetum is a European leader in digital services, supporting organizations as they navigate continuous technological change. The company helps clients accelerate their digital transformation through a broad portfolio that includes consulting, application services, digital engineering, cloud, cybersecurity, platforms, and infrastructure services

Job Description

The Application Production Support Engineer is a detail-oriented and technically skilled professional responsible for managing and supporting critical business applications in a production environment, with a strong focus on delivering high-quality IT services.

This role is responsible for ensuring the stability, availability, performance, and reliability of applications within the team's scope. The professional will also manage incident resolution in a timely manner, collaborating closely with internal stakeholders (Development, Infrastructure, and Architecture teams) as well as external partners and service providers to drive sustainable and long-term solutions.

Main Responsibilities

Application Stability & Availability (Expert)

  • Monitor, maintain, and support business-critical applications, ensuring high availability, stability, and optimal performance.
  • Actively participate in Incident and Problem Management processes, including:
    • Participation in Major Incident (P1/P2) Situation Rooms.
    • Root Cause Analysis (RCA) activities.
    • Identification of incident trends and recurring issues.
    • Contribution to permanent corrective actions and preventive measures.
  • Ensure compliance with ITIL governance practices, operational procedures, and established Service Level Agreements (SLAs).
  • Execute application deployments, releases, and change requests following ITIL and DevOps methodologies.
  • Proactively identify, analyze, and resolve technical issues to support uninterrupted business operations.
  • Participate in on-call support rotations and provide 24/7 support coverage for critical applications when required.

Technical Support & Collaboration (Expert)

  • Serve as a primary point of contact between Production Support and Development teams for troubleshooting and issue resolution.
  • Collaborate closely with Scrum and DevOps teams to design, deploy, maintain, and continuously improve application services.
  • Implement upgrades, patches, configuration changes, and new functionalities while minimizing business impact and ensuring service continuity.
  • Contribute to the continuous improvement of operational processes, automation, and platform reliability.

Documentation & Knowledge Sharing (Expert)

  • Create, maintain, and continuously update technical documentation, including operational procedures, system configurations, troubleshooting guides, and runbooks.
  • Promote knowledge sharing and best practices across global support teams to improve operational efficiency and service quality.
  • Support the development and maintenance of knowledge bases to facilitate issue resolution and onboarding activities.

Platform Monitoring & Observability (Expert)

  • Implement, maintain, and optimize monitoring and observability solutions across production environments.
  • Leverage platforms such as Dynatrace and other observability tools to ensure proactive monitoring and rapid incident detection.
  • Collaborate with Development teams and Centers of Expertise to define and enhance observability standards and monitoring strategies.
  • Promote an observability-first mindset, enabling early identification and resolution of potential service disruptions.
  • Continuously improve monitoring dashboards, alerting mechanisms, and operational metrics.

General Responsibilities

  • Complete all mandatory training required for the effective operation of the IT Production area and compliance with company policies.
  • Perform additional activities, when required, to support business objectives and ensure the proper functioning of the Center of Expertise.
  • Contribute to continuous improvement initiatives within the Production Support organization.

 

    Qualifications

    API, Application Servers & Kubernetes (Expert)

    • Strong knowledge of Java Application Servers, particularly Red Hat JBoss EAP.
    • Understanding of Java application troubleshooting, including:
      • Heap dump analysis.
      • Thread dump analysis.
      • Performance tuning and optimization.
    • Hands-on experience with OpenShift and Kubernetes-based platforms.
    • Knowledge of Cloud-native architectures and containerized environments.
    • Experience with API Gateway solutions such as Axway and Apigee.

    RHEL Linux Operating System (Expert)

    • Advanced administration and troubleshooting experience in Red Hat Enterprise Linux (RHEL) environments.

    Tooling & Automation (Expert)

    • Monitoring and observability platforms:
      • Dynatrace
      • Grafana
      • Prometheus
      • ELK Stack
      • Jaeger
    • CI/CD tools and practices:
      • GitLab
      • Jenkins
      • Argo CD
      • Nexus Repository (Sonatype)
    • Infrastructure and configuration management:
      • Ansible
      • Terraform

    Databases (Expert)

    • Strong knowledge of relational databases, including:
      • Microsoft SQL Server
      • PostgreSQL
    • Experience with database monitoring, performance analysis, and troubleshooting.

    Agile Methodologies

    Scrum Framework (Practitioner)

    • Experience working within Agile and Scrum environments, collaborating effectively with cross-functional teams.

    Additional Information

    English mandatory.

    Skills Required

    • Strong knowledge of Java application servers, particularly Red Hat JBoss EAP.
    • Understanding of Java application troubleshooting (heap dump and thread dump analysis, performance tuning).
    • Hands-on experience with OpenShift and Kubernetes-based platforms.
    • Knowledge of cloud-native architectures and containerized environments.
    • Experience with API Gateway solutions such as Axway and Apigee.
    • Advanced administration and troubleshooting experience in Red Hat Enterprise Linux (RHEL).
    • Experience with monitoring and observability platforms: Dynatrace, Grafana, Prometheus, ELK Stack, Jaeger.
    • Experience with CI/CD tools and practices: GitLab, Jenkins, Argo CD, Nexus Repository (Sonatype).
    • Experience with infrastructure and configuration management: Ansible, Terraform.
    • Strong knowledge of relational databases, including Microsoft SQL Server and PostgreSQL.
    • Experience working within Agile/Scrum environments.
    • English language proficiency (mandatory).
    Am I A Good Fit?
    beta
    Get Personalized Job Insights.
    Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

    The Company
    HQ: Saint-Ouen
    20,111 Employees

    What We Do

    Inetum is a European leader in digital services. Inetum’s team of 28,000 consultants and specialists strive every day to make a digital impact for businesses, public sector entities and society. Inetum’s solutions aim at contributing to its clients’ performance and innovation as well as the common good. Present in 19 countries with a dense network of sites, Inetum partners with major software publishers to meet the challenges of digital transformation with proximity and flexibility. Driven by its ambition for growth and scale, Inetum generated sales of 2.5 billion euros in 2023. Top Employer Europe 2024

    Similar Jobs

    Cloudflare Logo Cloudflare

    Business Development Representative

    Cloud • Information Technology • Security • Software • Cybersecurity
    Hybrid
    Lisbon, PRT
    4400 Employees
    38K-53K Annually

    Cloudflare Logo Cloudflare

    Business Development Representative

    Cloud • Information Technology • Security • Software • Cybersecurity
    Hybrid
    2 Locations
    4400 Employees
    38K-53K Annually

    Deepgram Logo Deepgram

    Account Executive

    Artificial Intelligence • Machine Learning • Natural Language Processing • Software • Conversational AI
    In-Office or Remote
    28 Locations
    150 Employees

    Datadog Logo Datadog

    Senior Software Engineer

    Artificial Intelligence • Cloud • Security • Software • Cybersecurity
    Easy Apply
    Remote or Hybrid
    Portugal
    6500 Employees

    Similar Companies Hiring

    Amplify Platform Thumbnail
    Fintech • Financial Services • Consulting • Cloud • Business Intelligence • Big Data Analytics
    Scottsdale, AZ
    62 Employees
    Standard Template Labs Thumbnail
    Artificial Intelligence • Information Technology • Software
    New York, NY
    25 Employees
    Golden Pet Brands Thumbnail
    Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
    El Segundo, California
    178 Employees

    Sign up now Access later

    Create Free Account

    Please log in or sign up to report this job.

    Create Free Account