All jobs → Partner boards
Site Reliability & DevOps Engineer (Mid-Level) -Splunk, PageDuty
None • Fairfax, VA, US • published 2026-09-29 02:55:15.960684
Apply on source siteDescription
**About the Role** We are seeking a hands\-on **Site Reliability \& DevOps Engineer** to support, automate, and optimize our production infrastructure, monitoring pipelines, and incident management workflows. In this role, you will help manage log aggregation, tune proactive alerting, automate cloud infrastructure, and assist in integrating AI\-assisted developer tooling into our daily operational workflows. **Key Responsibilities** * **Incident Response \& Reliability:** Support on\-call operations and incident lifecycle management using **PagerDuty** to streamline response times and reduce Mean Time to Resolution (MTTR). * **Observability \& Log Analytics:** Maintain centralized log aggregation, dashboards, and alerting rules across hybrid environments using **Splunk** and network monitoring tools (e.g., Zabbix, rsyslog). * **Infrastructure Automation:** Write and maintain Infrastructure as Code (IaC) using **Terraform** and **Ansible** to manage containerized and virtualized environments (Docker, Proxmox/KVM, AWS). * **Scripting \& Tooling:** Develop custom automation scripts, internal API integrations, and maintenance tooling using **Python**, **Bash/Shell**, or **Go**. * **AI ...
Get jobs like this daily in Telegram: @careeri_bot