Staff Site Reliability Engineer
Job Description
Senior Site Reliability Engineer (SRE) — Palo Alto Networks
Company: Palo Alto Networks
Designation: Senior Site Reliability Engineer
Department: Infrastructure & Cloud Operations
Experience: 5+ years
Qualification: Bachelor’s/Master’s degree in Computer Science, Information Technology, or related technical field
Work Mode: Primarily office-based, with flexibility where needed
Industry: Cybersecurity / Cloud Infrastructure / IT Operations
Key Responsibilities
Implement and support Linux infrastructure as code.
Provision, configure and maintain resilient hybrid-cloud architectures.
Manage Linux infrastructure CI/CD platforms and automation frameworks.
Handle scalability, capacity planning, redundancy and resiliency.
Maintain availability and performance SLAs.
Build and operate compute infrastructure supporting thousands of VMs and Kubernetes clusters.
Develop scripts and tools to automate routine operational tasks.
Design monitoring, alerting and trend-analysis solutions.
Support security implementations and audits.
Prepare documentation for deployment, operations and DR/BCP.
Plan maintenance windows and technical change requests.
Participate in on-call support and Root Cause Analysis (RCA).
Implement automated processes to improve operational efficiency.
Collaborate with Network, Compute, Security, Database and Application teams.
Work with globally distributed teams across multiple time zones.
Required Skills
Linux infrastructure administration
Docker & Kubernetes
AWS / GCP or other cloud platforms
Infrastructure as Code:
Terraform
Ansible
Git
Puppet
Python / Shell / Bash scripting
CI/CD — Jenkins, CircleCI, etc.
Monitoring & Observability
MELT — Metrics, Logs, Events & Traces
Network and security technologies
API development, optimization and security
Automation and infrastructure management
Cross-functional collaboration
Additional / Preferred Skills
AIOps
Machine Learning / AI for cloud infrastructure and IT operations
Self-healing infrastructure
Big Data and data analytics
Enterprise Business Applications
ITSM frameworks and tools
Category: Site Reliability Engineering / DevOps / Cloud Infrastructure / Linux / Kubernetes / Automation / Observability / Cybersecurity