%term

Gartner: tips for improving reliability

Jun 6, 2022 By Andre Newman In Gremlin

In their report titled “IT Resilience — 7 Tips for Improving Reliability, Tolerability and Disaster Recovery”, Gartner presents seven strategies for improving the resilience posture of your critical systems. These recommendations range from how to get started, to identifying IT hazards and risks to reliability, to capturing metrics and translating them into business value. In this blog, we’ll take a high-level look at the report and summarize some of its key findings.

Read Post

Gremlin

Read more about Gartner: tips for improving reliability

Podcast: Break Things on Purpose | KubeCon, Kindness, and Legos with Michael Chenetz

May 31, 2022 By Jason Yee In Gremlin

In this episode, we chat with Cisco’s head of developer content, community, and events, Michael Chenetz. We discuss everything from KubeCon to kindness and Legos! Michael delves into some of the main themes he heard from creators at KubeCon, and we discuss methods for increasing adoption of new concepts in your organization. We have a conversation about attending live conferences, COVID protocol, and COVID shaming, and then we talk about how Legos can be used in talks to demonstrate concepts.

Read Post

Gremlin

Read more about Podcast: Break Things on Purpose | KubeCon, Kindness, and Legos with Michael Chenetz

Site Reliability Chats (May 18, 2022)

May 18, 2022 By Gremlin In Gremlin

View Video

Gremlin

Read more about Site Reliability Chats (May 18, 2022)

Podcast: Break Things on Purpose | Dan Isla: Astronomical Reliability

May 17, 2022 By Jason Yee In Gremlin

It’s time to shoot for the stars with Dan Isla, VP of Product at itopia, to talk about everything from astronomical importance of reliability to time zones on Mars. Dan’s trajectory has been a propulsion of jobs bordering on the science fiction, with a history at NASA, modernizing cloud computing for them, and loads more. Dan discusses the finite room for risk and failure in space travel with an anecdote from his work on Curiosity.

Read Post

Gremlin

Read more about Podcast: Break Things on Purpose | Dan Isla: Astronomical Reliability

Site Reliability Chats (May 11, 2022)

May 11, 2022 By Gremlin In Gremlin

View Video

Gremlin

Read more about Site Reliability Chats (May 11, 2022)

Introduction to GameDay webinar

May 10, 2022 By Gremlin In Gremlin

Learn all about Gremlin's GameDay feature in this webinar presented by Sydney Lesser and Andre Newman. GameDays are organized team events to proactively improve reliability using Chaos Engineering principles. Gremlin makes it easier than ever to prepare, execute, and learn from them. Increase your system’s reliability with safe, secure, and simple GameDays.

View Video

Gremlin

Read more about Introduction to GameDay webinar

How to run a GameDay using Gremlin

May 10, 2022 By Gremlin In Gremlin

Learn how to run a GameDay in Gremlin. This video walks you through creating a GameDay, adding and running Scenarios, recording your observations, and linking to Jira in the Gremlin web app.

View Video

Gremlin

Read more about How to run a GameDay using Gremlin

How Gremlin runs a GameDay

May 10, 2022 By Sydney Lesser In Gremlin

You might be familiar with GameDays at this point. From watching our Introduction to GameDay webinar, viewing our Demo video, and reading our tutorial, you’ve probably learned that GameDays were created with the goal of increasing reliability by purposely creating major failures on a regular basis. Better yet, perhaps your own team has run a GameDay and learned something new about their services’ behavior during failure scenarios.

Read Post

Gremlin

Read more about How Gremlin runs a GameDay

Site Reliability Chats (May 4, 2022)

May 4, 2022 By Gremlin In Gremlin

View Video

Gremlin

Read more about Site Reliability Chats (May 4, 2022)

Podcast: Break Things on Purpose | Natalie Conklin: Learning to Embrace Change

May 3, 2022 By Julie Gunderson In Gremlin

Natalie Conklin, tamer of chaos and Head of Engineering here at Gremlin, joins us to talk about embracing change, working alongside each other, and building more reliable systems. Natalie has a talk coming up at DevOpsDays Boise which she has titled “Embracing Change Fearlessly.” Her talk is oriented around enabling teams to take calculated risks and having the guts to take those risks. Natalie spent time working in India, which helped solidify her “fearlessly” philosophy.

Read Post

Gremlin

Read more about Podcast: Break Things on Purpose | Natalie Conklin: Learning to Embrace Change

Operations | Monitoring | ITSM | DevOps | Cloud

Gartner: tips for improving reliability

Podcast: Break Things on Purpose | KubeCon, Kindness, and Legos with Michael Chenetz

Site Reliability Chats (May 18, 2022)

Podcast: Break Things on Purpose | Dan Isla: Astronomical Reliability

Site Reliability Chats (May 11, 2022)

Introduction to GameDay webinar

How to run a GameDay using Gremlin

How Gremlin runs a GameDay

Site Reliability Chats (May 4, 2022)

Podcast: Break Things on Purpose | Natalie Conklin: Learning to Embrace Change

Monthly Archive

Follow Us