Description
Summary:
Seeking a DevOps/SRE to support development, operation, and maintenance of large-scale distributed systems, focusing on infrastructure reliability, Kubernetes operations, and continuous improvement.
Highlights:
1. Support large-scale distributed systems on Linux and Open Source technologies
2. Focus on infrastructure reliability, Kubernetes, automation, and observability
3. Contribute to scalability, performance, and high availability initiatives
**Description du poste**
We are looking for a **DevOps / Site Reliability Engineer (SRE)** to support the development, operation and maintenance of large\-scale distributed systems running on Linux and Open Source technologies. This role focuses on infrastructure reliability, Kubernetes operations, automation, observability and the continuous improvement of critical production environments.
**Main Responsibilities**
* Administer and operate Linux\-based distributed systems and platforms.
* Manage and support Kubernetes clusters in on\-premises environments.
* Implement and maintain DevOps and automation tools.
* Monitor, troubleshoot and optimize critical production systems.
* Support CI/CD implementation and operational practices.
* Administer relational and non\-relational database platforms.
* Maintain observability, logging and monitoring solutions.
* Contribute to infrastructure scalability, performance and high availability initiatives.
* Provide operational support and participate in on\-call activities.
**Profil**
* Minimum 5 years of experience in Linux systems administration and Open Source environments.
* Experience administering Kubernetes clusters.
* Experience with MySQL and PostgreSQL.
* Experience with Apache and Nginx environments.
* Experience with CI/CD tools, including Jenkins.
* Experience with email platforms and protocols (IMAP, SMTP, POP).
* Experience with Elasticsearch, Kibana, Prometheus and Grafana.
* Knowledge of Redis, MongoDB and other non\-relational databases.
* Experience working with highly available, large\-scale production environments.
* Bash and Python scripting skills.
* Knowledge of Docker, Kubernetes, networking and TCP/IP protocols.
* Availability to participate in on\-call support rotations.
* Experience using GitHub Copilot Business or equivalent tools (subject to validation).
**Additional Valued Experience**
* RabbitMQ or Kafka.
* Java.
* Puppet, Ansible and ArgoCD.
* GitLab CI.
* Azure cloud and DevOps technologies.
* Golang.
* Rancher.
**À propos d'Inetum**
Inetum is a **European leader in digital services**, supporting organizations as they navigate continuous technological change. The company helps clients accelerate their digital transformation through a broad portfolio that includes **consulting, application services, digital engineering, cloud, cybersecurity, platforms, and infrastructure services.**
**Country**
Portugal
**Location**
Lisbon
**Contract Type**
Fixed\-term contract