Jennifer Petoff
The Reliable PgM, Co-Editor of the SRE Book
Lisbon, Portugal
Actions
Jennifer Petoff is a globally recognized expert on SRE education. She is the co-editor of the bestselling Site Reliability Engineering book and the lead author of Training Site Reliability Engineers: What Your Organization Needs to Create a Learning Program. Known as The Reliable PgM, she is an advocate for applying SRE principles to a wide range of program management situations. Jennifer is an avid public speaker and has given talks, keynote presentations and participated in panel discussions at DevOps, SRE, and other industry conferences in 16 countries (and counting) on four continents. Jennifer worked at Google for 19 years (most recently as a program management director) after spending eight years in the chemical industry. She holds a PhD in chemistry from Stanford University, along with a BS in chemistry and a BA in psychology from the University of Rochester in New York. Jennifer and her husband Scott are avid travelers and have lived in the US, Ireland, and now Portugal. Jennifer loves travel writing and door photography, both of which you can find on Sidewalk Safari.
Area of Expertise
Topics
New Grads Becoming New SREs: Catalyzing a “Circle of Life” in Ireland
Site Reliability Engineering principles, best practices, and culture do not feature systematically in the undergraduate curriculum around the world. Nor do principles of non-abstract large system design. Despite this, students can be taught (and learn through experience) to be great SREs upon graduation.
This talk will equip SRE hiring managers with creative ways to build a pipeline of talent. We’ll share techniques that we’ve found to be effective in super-charging our SRE hiring pipeline from universities in Ireland.
Site Reliability Engineering to build high performance software and teams
Site Reliability Engineering (SRE) is a discipline founded at Google that is now widely practiced across the Tech industry. SRE represents a set of principles and practices that applies aspects of software engineering to IT infrastructure and operations. In this talk, we will discuss the key principles and practices of SRE, and how they can be used to build high performance software and teams. We’ll explore insights from the State of DevOps Report and how SRE can help foster the type of generative organizational culture that is a hallmark of high performing organizations.
Swim Don’t Sink: Why Training Matters to an SRE Practice
Do you offer training to the engineers in your organization or do you throw them off the deep end to “sink or swim”? Providing training and education is universally important to set team members up for success in your organization and is critical for establishing a thriving Site Reliability Engineering (SRE) or DevOps practice and culture in the first place.
The specific training needs of each engineer varies depending on several factors including:
-The maturity of your organization in adopting DevOps / SRE principles, practices, and culture
-The knowledge those individuals have about your organization and infrastructure
-The experience of the individuals being trained, both in terms of technical skill and familiarity with the SRE / DevOps model
This talk will explore the business case for training, the trade-offs between cost and effectiveness, and best practices for training design and deployment depending on where your organization lies on the spectrum of size and maturity.
Learn why training is not about unleashing a fire hose of information upon unsuspecting engineers but about giving those engineers the confidence to run production systems at scale.
Site Reliability Engineering: Anti-patterns in Everyday Life and What They Teach Us
Real world experience and things that go wrong are two of life’s best teachers. This talk will explore key elements of scalable large-system design and Site Reliability Engineering (SRE) principles* through anti-patterns encountered in real life. Find out what lessons can be gleaned from watching the dynamics in a crowded cafe or dealing with a security issue during a hotel stay. Learn about fundamental site reliability engineering principles and practices including:
-Avoiding cascading failures
-Not feeding the machines with human toil
-Writing blameless postmortems
-Engineering solutions to eliminate classes of errors rather than implementing point fixes
These principles will be framed through a lens of the suboptimal while demonstrating the impact of SRE anti-patterns on user trust.
* SRE is often thought of as a specific implementation of the DevOps interface.
How to Run Smarter in Production: Getting Started With Site Reliability Engineering
Site Reliability Engineering and the DevOps movement share a similar set of challenges but addresses each in a different way. SRE got its start at Google in 2003 and according to Ben Treynor, VP of 24/7 Operations: ”SRE is what happens when you ask a software engineer to design an operations team”. In 2016, Google published a book about Site Reliability Engineering principles, practices and organizational constructs.
The practice of Site Reliability Engineering at Google encompasses more than just managing production systems and responding to emergencies. Applying software engineering in a principled way to operations allows SRE to holistically address the reliability of software applications across the product lifecycle.
Implementing SRE in an organization requires a commitment to supporting some core principles and a fundamental culture shift -SRE needs Service Level Objectives, with consequences.
-SREs have time to make tomorrow better than today.
-SRE teams have the ability to regulate their workload.
-SREs and the organization’s leaders remove the word ‘blame’ from their vocabulary.
This talk will highlight key SRE principles and how they map to recognized DevOps focus areas. We’ll also discuss how any organization can adopt SRE, and how our recent experience of working with our customers on implementing SRE practices has shown these principles will work across a range of organizations of different types and sizes.
Enterprise Technology Leadership Summit Virtual 2024 Sessionize Event
DevOpsDays Zurich 2024 Sessionize Event
Incontro DevOps Italia (IDI) 2024 Sessionize Event
DevOps Enterprise Summit Amsterdam 2023 Sessionize Event
WorldFestival 2021 Sessionize Event
2020 All Day DevOps Sessionize Event
GDG DevFest UK & Ireland 2020 Sessionize Event
All Day DevOps: Spring Break Edition Sessionize Event
2019 All Day DevOps Sessionize Event
Please note that Sessionize is not responsible for the accuracy or validity of the data provided by speakers. If you suspect this profile to be fake or spam, please let us know.
Jump to top