AgileEngine is seeking a Senior Site Reliability Engineer to provide core administration and stability for enterprise systems, including on‑premise and SaaS environments. You will manage Kubernetes clusters, enable observability with Snowflake/OpenTelemetry, and participate in on‑call rotations and RCA.
You will automate infrastructure tasks with Bash, Python, or Go, contributing to reliability across ESM and ECP platforms, while collaborating with integral teams on modern projects.