CloudOps Engineer
Location: Lisbon, Portugal (100% Onsite)
Experience: 5+ Years
Language Requirement: Fluent English (Written & Spoken)
About the Role
We are seeking an experienced CloudOps Engineer to join a fast-paced, technology-driven environment where reliability, automation, security, and scalability are key priorities. This is a hands-on role focused on building and operating modern cloud infrastructure, automating delivery processes, improving platform reliability, and enabling engineering teams to move faster with confidence.
You will play a critical role in designing and maintaining cloud-native platforms, implementing infrastructure as code, enhancing observability, and ensuring operational excellence across complex distributed systems.
Key Responsibilities
- Build, manage, and optimize cloud environments using Infrastructure as Code (IaC) and reusable automation frameworks.
- Design and maintain CI/CD pipelines with automated testing, validation, security controls, and deployment strategies.
- Implement and improve observability across logs, metrics, traces, and monitoring platforms.
- Define and maintain reliability standards, including SLOs, error budgets, incident response processes, and operational best practices.
- Establish secure-by-default environments, infrastructure, and deployment patterns.
- Drive automation initiatives to eliminate manual operational tasks and improve platform efficiency.
- Optimize cloud performance, scalability, and cost through continuous monitoring and analysis.
- Collaborate closely with engineering, security, data, and platform teams to ensure reliable and secure service delivery.
- Support incident management, root cause analysis, and continuous improvement initiatives.
Technology Environment
You will work with technologies and practices including:
- Infrastructure as Code (Terraform, Bicep, Pulumi, Flux, and policy-driven automation)
- Cloud platforms, with a strong preference for Microsoft Azure
- Multi-stage CI/CD pipelines and deployment automation
- Observability and monitoring solutions
- Distributed systems supporting data, analytics, streaming, and AI workloads
- Security technologies including secrets management, image signing, and zero-trust principles
- Autoscaling, performance optimization, and FinOps practices
Required Qualifications
- Experience in Cloud Operations, DevOps, Site Reliability Engineering, or a related field.
- Fluent English communication skills (written and spoken).
- Strong experience with cloud-native environments, preferably Microsoft Azure.
- Proven experience building and operating large-scale automated infrastructure and deployment pipelines.
- Deep understanding of observability, monitoring, logging, telemetry, and reliability engineering principles.
- Experience troubleshooting and supporting distributed systems in production environments.
- Strong automation mindset with expertise in scripting and infrastructure management.
- Knowledge of security best practices for cloud environments and CI/CD ecosystems.
- Ability to thrive in a fast-moving environment and manage multiple priorities effectively.
- Excellent problem-solving and stakeholder communication skills.