Stellendetails
Revolutionierender Schutz.
Definieren Sie die Zukunft der Cybersicherheit.
Site Reliability Engineer
Our Mission
At Palo Alto Networks®, we’re united by a shared mission—to protect our digital way of life. We thrive at the intersection of innovation and impact, solving real-world problems with cutting-edge technology and bold thinking. Here, everyone has a voice, and every idea counts. If you’re ready to do the most meaningful work of your career alongside people who are just as passionate as you are, you’re in the right place.
Who We Are
In order to be the cybersecurity partner of choice, we must trailblaze the path and shape the future of our industry. This is something our employees work at each day and is defined by our values: Disruption, Collaboration, Execution, Integrity, and Inclusion. We weave AI into the fabric of everything we do and use it to augment the impact every individual can have. If you are passionate about solving real-world problems and ideating beside the best and the brightest, we invite you to join us!
We believe collaboration thrives in person. That’s why most of our teams work from the office full time, with flexibility when it’s needed. This model supports real-time problem-solving, stronger relationships, and the kind of precision that drives great outcomes.Job Summary
Job Summary
We are seeking a highly skilled Site Reliability Engineer (SRE) to ensure the reliability, scalability, and performance of our cloud-based SaaS services and AWS infrastructure. In this pivotal role, you will drive and improve incident management processes, collaborating closely with development and operations teams to maintain resilient systems. Your expertise will be crucial in monitoring system health, maintaining fault-tolerant distributed systems, and optimizing performance. Join our team to tackle complex challenges and contribute to the stability of world-class security solutions.
Key Responsibilities
- Drive incident response processes and troubleshoot complex issues to ensure timely resolution of outages.
- Establish best practices for monitoring, logging, and alerting using tools like Datadog and Site24x7.
- Build essential tooling and automation to improve system reliability and enable automated issue remediation.
- Participate in a 24x7 on-call rotation to ensure continuous system availability and support.
- Create and maintain clear documentation for infrastructure, processes, and incident management protocols.
- Automate provisioning, configuration, and deployment using Infrastructure as Code (IaC) tools like Terraform and Ansible.
- Continuously optimize system performance by identifying bottlenecks and improving scalability and efficiency.
- Identify and implement strategies to optimize cloud costs while maintaining system performance and reliability.
- Implement security best practices to protect infrastructure and data from vulnerabilities and threats.
- Collaborate with cross-functional teams to understand requirements and provide technical guidance.
- Implement AI-based automations and solutions to improve productivity and share best practices.
Qualifications
Required Qualifications
- 0-2 years of experience as a Site Reliability Engineer or Cloud Engineer.
- Strong proficiency in AWS cloud services (e.g., EC2, S3, VPC, RDS, EKS, ECS, CloudFormation).
- Strong logical, analytical, and problem-solving skills.
- Strong communication skills and the ability to work in a 24x7 shift rotation.
- Strong scripting skills in languages such as Python, PowerShell, or Shell.
- Understanding of Infrastructure as Code (IaC) tools like Terraform or Ansible.
- Knowledge of containerization (Docker) and orchestration platforms (Kubernetes).
- Experience with CI/CD pipelines and tools like Jenkins or GitHub.
- Experience with monitoring and alerting tools like CloudWatch, Datadog, or Grafana.
- Experience documenting Standard Operating Procedures (SOPs) and Root Cause Analyses (RCAs).
- Understanding of security best practices and compliance standards.
Preferred Qualifications
- AWS Certification.
- Experience with AWX Tower for Ansible automation.
- Security Certification.
- Familiarity with AI-assisted software development and productivity tools.
Our Commitment
We’re trailblazers that dream big, take risks, and challenge cybersecurity’s status quo. It’s simple: we can’t accomplish our mission without diverse teams innovating, together.
We are committed to providing reasonable accommodations for all qualified individuals with a disability. If you require assistance or accommodation due to a disability or special need, please contact us at accommodations@paloaltonetworks.com.
Palo Alto Networks is an equal opportunity employer. We celebrate diversity in our workplace, and all qualified applicants will receive consideration for employment without regard to age, ancestry, color, family or medical care leave, gender identity or expression, genetic information, marital status, medical condition, national origin, physical or mental disability, political affiliation, protected veteran status, race, religion, sex (including pregnancy), sexual orientation, or other legally protected characteristics.
All your information will be kept confidential according to EEO guidelines.
Is role eligible for Immigration Sponsorship? No. Please note that we will not sponsor applicants for work visas for this position.Mehr zu Palo Alto Networks
-
Eine SaaS-Unternehmensgeschichte.
So hat Palo Alto Networks kritische SaaS-Apps mit SaaS Security Posture Management gesichert.
-
Unsere Kultur
Wegweisend in einer globalen Gemeinschaft – von der Vision zur Tat
-
Berufseinsteiger & Nachwuchsprogramme
Our early-in-career programs will train you to be a part of the next generation of cybersecurity talent.
Keine kürzlich angesehenen Jobs
Keine kürzlich angesehenen Jobs