Working at Infobip means being part of something truly integral. With 75+ offices across six continents, we’re not just building technology — we’re shaping how more than 80% of the world connects and communicates.
As employees, we take pride in contributing to the world’s largest and only full-stack cloud communication platform. But it’s not just what we do, it’s how we do it: with curiosity, passion, and a whole lot of collaboration.
We operate with an AI-first mindset, embedding intelligent tools into our daily workflows to work smarter and more efficiently. Every role here benefits from and contributes to this approach.
If you're looking for meaningful work and challenges that grow you in a culture where people show up with purpose, this is your opportunity.
Let’s build what’s next, together.
What This Role Is All About
As a Reliability Operations Engineer you will ensure the stability, reliability, and continuous improvement of our platform. You will play a key role in incident management, monitoring, automation within your team’s scope.
This role combines operational excellence, problem-solving, and engineering ownership.
What You’ll Do
- Create, respond to, and continuously improve platform alerts and runbooks
- Actively monitor the platform, identify issues, and triage incidents
- Perform impact assessments and communicate incident summaries clearly
- Escalate incidents to the correct owner teams and act as Incident Leader when required
- Execute mitigation actions to minimize impact and restore service
- Write, test, secure, and maintain well-documented scripts and automation
- Work autonomously on complex technical tasks and initiatives
- Solve challenging technical problems in collaboration with senior engineers
- Ensure stable and reliable service delivery within the team’s scope
- Provide and receive constructive feedback to continuously improve performance
What Makes You a Strong Fit
- Strong understanding of monitoring, alerting, and incident mana