CLOUD SUPPORT INFRASTRUCTURE EXPERT (HYBRID)
iTRTech Group · Lisbon, Portugal
You apply off-site, with the employer or the job board. I never handle applications.
CLOUD SUPPORT INFRASTRUCTURE EXPERT (HYBRID LISBON)
Portuguese company hires for hybrid position
Location: Lisbon, Portugal
- ️ Only candidates already based in Portugal will be considered
Work Model: Hybrid — remote work is available
️ Language Requirements: English C1 — mandatory
Seniority: Senior (5+ years)
Sector: Banking
Rate Between €4500 - 4800 RV / €2600 - 2900 CTI
- ️ Instructions: Please send your CV in English and make sure to include all skills and experience that match the requirements of the opportunity. This will significantly increase your chances of success
About the Opportunity
We are looking for an experienced Cloud Support Infrastructure Expert to join an Application Production Support team responsible for business-critical Wealth Management applications and data platforms.
You will play an essential role in maintaining the stability, availability and performance of production environments. Your responsibilities will include incident management, application support, infrastructure automation, monitoring, disaster recovery exercises and continuous service improvement.
This is an excellent opportunity for a hands-on infrastructure specialist who enjoys solving complex production issues, automating operational processes and working in an international banking environment.
Main Responsibilities
- Provide application production support for Wealth Management data platforms, applications and tools.
- Manage incidents, service requests and other application support activities.
- Analyse system, middleware and application logs to identify root causes and potential issues.
- Monitor services and produce reports regarding performance, quality and availability.
- Build and maintain Ansible playbooks and Terraform scripts for provisioning, configuration and deployments.
- Support cloud and containerised environments using Kubernetes and Docker.
- Train and assist other operations professionals in the use of Ansible, Docker and Kubernetes.
- Escalate operational risks, critical incidents and management issues in a timely manner.
- Ensure incidents and service requests are handled according to established service-level objectives.
- Create, review and maintain application support documentation and knowledge articles.
- Participate in shifts, on-call rotations and 24/7 production support activities.
- Support disaster recovery and live-play exercises.
- Participate in and validate knowledge-transfer sessions delivered by application development teams.
- Work closely with development and production support teams across different locations.
- Contribute to reducing Mean Time to Detect (MTTD) and Mean Time to Resolve (MTTR).
- Ensure production-readiness checks, monitoring standards and compliance requirements are followed.
Mandatory Technical Requirements
- Strong knowledge of Linux platforms, command-line tools and shell scripting.
- Proficiency in at least one scripting language, preferably Python or Bash.
- Hands-on experience with infrastructure automation using Ansible and Terraform.
- Practical experience with Kubernetes and Docker.
- Experience with at least one public cloud platform: IBM Cloud, AWS, Microsoft Azure or Google Cloud Platform.
- Knowledge of continuous integration tools such as Jenkins, Travis CI or Concourse.
- Experience analysing infrastructure, middleware and application logs.
- Understanding of production monitoring, incident management and troubleshooting.
- Knowledge of ITIL practices.
- Ability to work in shifts and participate in on-call rotations, including weekends and public holidays.
- C1-level English communication skills.
Qualifications and Experience
- Master’s degree in Computer Science, Computer Engineering or a related field.
- At least five years of experience in IT Operations, Infrastructure, Cloud Engineering, DevOps, Site Reliability Engineering or Software Engineering.
- At least two years of hands-on experience with a cloud provider.
- At least two years of practical experience with Ansible and Docker.
- Previous L1/L2 application or infrastructure support experience, preferably within banking or another regulated environment.
- Understanding of core banking operations is highly valued.
- Experience with PL/SQL and Dynatrace is strongly preferred.
Relevant Certifications
The following certifications are highly valued:
- Red Hat Enterprise Linux Automation with Ansible — RH294.
- Certified Kubernetes Application Developer — CKAD.
- Developing Advanced Automation with Red Hat Ansible Automation Platform — DO374.
- HashiCorp Certified: Terraform Associate — optional.
- HashiCorp Certified: Vault Associate — optional.
Soft Skills
- Strong analytical and critical-thinking abilities.
- Structured and practical approach to problem-solving.
- Excellent written and verbal communication skills.
- Proactive attitude and strong sense of ownership.
- Persistence when investigating complex production issues.
- Ability to collaborate effectively with global technical teams.
- Commitment to continuous learning and technical development.
- Strong awareness of operational, regulatory and compliance risks.
The Ideal Profile
The ideal candidate combines strong Linux and cloud infrastructure expertise with practical production support experience. You are comfortable investigating critical incidents, reading complex logs and automating repetitive operational tasks with Ansible, Terraform, Python or Bash.
You understand how Kubernetes, Docker, CI/CD tools and public cloud services work together in highly available environments. Previous experience in banking, Wealth Management or another regulated sector will be particularly valuable.
You are also available to participate in a 24/7 support structure and remain calm, organised and communicative during high-impact incidents.
Questions for Candidates
Please include answers to the following questions with your application:
- Are you currently based in Portugal and available to work in a hybrid model in Lisbon?
- What is your English proficiency level?
- How many years of experience do you have in IT Operations, Cloud Engineering, DevOps or Production Support?
- Which cloud platforms have you worked with: IBM Cloud, AWS, Azure or Google Cloud?
- How many years of hands-on experience do you have with Ansible, Terraform, Kubernetes and Docker?
- Are you proficient in Python, Bash or another scripting language?
- Have you provided L1/L2 application or infrastructure support in a banking or regulated environment?
- Do you have practical experience with Linux, PL/SQL and Dynatrace?
- Have you created or maintained Ansible playbooks and Terraform scripts?
- Do you hold any Red Hat, Kubernetes or HashiCorp certifications?
- Are you available for shifts, on-call rotations, weekends and public holidays?
- Which contract model do you prefer: RV/B2B or CTI?
- What is your expected monthly rate or salary?
- What is your availability to start?
CV Keywords
Cloud Support Infrastructure Expert, Application Production Support, APS, IT Operations, Cloud Engineering, DevOps, Site Reliability Engineering, SRE, Linux, Red Hat Enterprise Linux, RHEL, Linux CLI, Shell Scripting, Bash, Python, PL/SQL, Dynatrace, Ansible, Ansible Playbooks, Ansible Automation Platform, Terraform, HashiCorp Terraform, HashiCorp Vault, Kubernetes, K8s, Docker, IBM Cloud, AWS, Microsoft Azure, Google Cloud Platform, GCP, Public Cloud, Infrastructure Automation, Infrastructure as Code, IaC, Jenkins, Travis CI, Concourse, Continuous Integration, CI/CD, Application Support, Production Support, L1 Support, L2 Support, Incident Management, Service Requests, Troubleshooting, Log Analysis, Middleware, Monitoring, Service Availability, ITIL, SLA, SLO, MTTD, MTTR, Disaster Recovery, DR, Live Play, On-Call, 24/7 Support, Wealth Management, Banking, Financial Services, Compliance, Risk Mitigation, RH294, CKAD, DO374
#CI - PROC26354