AgileEngine is seeking a Senior Site Reliability Engineer to provide core system administration and ensure operational stability across on-premise and SaaS-hosted environments, with a strong focus on Kubernetes, monitoring, and observability.
You will participate in on-call rotations, incident response, and RCA, automate infrastructure tasks with Bash, Python, or Go, and collaborate with integral teams to keep services healthy and scalable.