CareeriAll jobsEmail digestTelegram bot

All jobs → Partner boards

Site Reliability & DevOps Engineer (Mid-Level) -Splunk, PageDuty

None • Fairfax, VA, US • published 2026-09-29 02:55:15.960684

Apply on source site

Description

**About the Role** We are seeking a hands\-on **Site Reliability \& DevOps Engineer** to support, automate, and optimize our production infrastructure, monitoring pipelines, and incident management workflows. In this role, you will help manage log aggregation, tune proactive alerting, automate cloud infrastructure, and assist in integrating AI\-assisted developer tooling into our daily operational workflows. **Key Responsibilities** * **Incident Response \& Reliability:** Support on\-call operations and incident lifecycle management using **PagerDuty** to streamline response times and reduce Mean Time to Resolution (MTTR). * **Observability \& Log Analytics:** Maintain centralized log aggregation, dashboards, and alerting rules across hybrid environments using **Splunk** and network monitoring tools (e.g., Zabbix, rsyslog). * **Infrastructure Automation:** Write and maintain Infrastructure as Code (IaC) using **Terraform** and **Ansible** to manage containerized and virtualized environments (Docker, Proxmox/KVM, AWS). * **Scripting \& Tooling:** Develop custom automation scripts, internal API integrations, and maintenance tooling using **Python**, **Bash/Shell**, or **Go**. * **AI ...


Get jobs like this daily in Telegram: @careeri_bot