Description
Job Summary:
We are seeking a detail-oriented professional to ensure operational stability of systems, monitor system performance, manage incidents, and support application deployment and configuration.
Key Highlights:
1. Ensure operational stability of platforms
2. Participate in load and fault tolerance testing
3. Willingness to continuously evolve and learn
Do you enjoy ensuring systems run without failure and pay close attention to detail? This challenge is for you!
**What do you need?**
* Prior experience in system operations or monitoring roles
* Familiarity with monitoring tools and service management platforms
* Ability to interpret logs and identify error messages
* Strong analytical skills, incident escalation capability, and report writing proficiency
* Orientation toward following procedures and standards in production environments
**Your Role?**
You will be a key contributor to platform operational stability:
* Monitor dashboards and identify error alerts or communication failures
* Execute operational routines (start/stop services, check logs/records)
* Log and track incidents in service management platforms, escalating as needed
* Support application deployment and configuration, validating correct functionality after changes
* Participate in load and fault tolerance testing alongside the infrastructure team
Create and update simple technical documentation to support process standardization
**What do we value?**
* Attention to detail and rigor in procedure execution
* Proactivity and ability to suggest improvements to monitoring/alarm rules
* Strong communication skills and collaborative mindset
* Willingness to continuously evolve and learn in a technology-driven environment
Do you want to join this team and ensure critical systems never go down?
Apply now and become a true **Nimber**!