%term

Office Hours: How to test serverless applications using Failure Flags

Oct 10, 2024 By Gremlin In Gremlin

Part of the Gremlin Office Hours series: A monthly deep dive with Gremlin experts. Serverless applications are ideal for deploying scalable applications without having to manage infrastructure. However, this also makes it difficult to test their reliability. It’s easy to simulate a network outage or latency when you have direct access to the host that your software’s running on. What do you do when you only have control over the code?

View Video

Gremlin

Read more about Office Hours: How to test serverless applications using Failure Flags

How Visa Cross Border Solutions Reduces Outages by Testing System Resilience in Their SDLC

Oct 7, 2024 By Gremlin In Gremlin

For global financial services companies, reliability must be built-in and validated before and after shipping to production. Resilience testing is crucial for verifying the reliability of your applications under real-world conditions. But ad-hoc testing and exploratory experiments aren't sufficient: you need to run automated, standardized tests at global scale.

View Video

Gremlin

Read more about How Visa Cross Border Solutions Reduces Outages by Testing System Resilience in Their SDLC

Interpreting your reliability test results

Sep 19, 2024 By Andre Newman In Gremlin

Gremlin’s default suite of reliability tests analyzes critical functions of modern services: scalability, redundancy, and resilience to dependency failures. Services that pass this suite of tests can be trusted to remain available during unexpected incidents. But what happens when a service fails a test? How do you take failed test results and turn them into actionable insights? This blog aims to answer that question.

Read Post

Gremlin

Read more about Interpreting your reliability test results

What's Chaos Monkey? Its Role in Modern Testing

Sep 17, 2024 By Muhammad Raza In Splunk

Chaos Monkey is an open-source tool. Its primary use is to check system reliability against random instance failures. Chaos Monkey follows the testing concept of chaos engineering, which prepares networked systems for resilience against random and unpredictable chaotic conditions. Let’s take a deeper look.

Read Post

Splunk

Read more about What's Chaos Monkey? Its Role in Modern Testing

Office Hours: Get better reliability on AWS with our new release

Sep 12, 2024 By Gremlin In Gremlin

Part of the Gremlin Office Hours series: A monthly deep dive with Gremlin experts. Cloud platforms make it easier than ever to deploy massively scalable, distributed workloads, but this is a double-edged sword. There are reliability challenges unique to the cloud that didn’t exist before. Failed migrations, recurring incidents, and reliability toil take their toll.

View Video

Gremlin

Read more about Office Hours: Get better reliability on AWS with our new release

Release Roundup August 2024

Sep 9, 2024 By Andre Newman In Gremlin

Over the past year, the Gremlin team has focused on giving you more tools to adapt Gremlin to your organization’s reliability needs. We started with customizable reliability tests, and now, we’ve released customizable role-based access controls (RBAC). We’ve also made it easier to target specific availability zones when running Failure Flags experiments, and to run experiments behind a proxy. Keep reading to learn more! ‍

Read Post

Gremlin

Read more about Release Roundup August 2024

Reliability recommendations when adopting Kubernetes

Sep 3, 2024 By Andre Newman In Gremlin

Kubernetes just celebrated its tenth birthday. That’s 10 years of microservices, containers, service meshes, and many other paradigms that are now common to many developers’ toolkits.

Read Post

Gremlin

Read more about Reliability recommendations when adopting Kubernetes

How to verify, document, and prove compliance with Gremlin

Aug 29, 2024 By Gavin Cahill In Gremlin

Resilient and reliable IT systems have become a minimum requirement for modern businesses—a fact driven home by any number of high-profile outages over the past few years. Unfortunately, when those outages are in the financial sector, it can have far-reaching and incredibly damaging results.

Read Post

Gremlin

Read more about How to verify, document, and prove compliance with Gremlin

Achieving SLO Success with Golden Signals and Reliability Testing

Aug 28, 2024 By Gremlin In Gremlin

The four Golden Signals are an easy and effective way to measure the most important aspects of a system, and when paired with a reliability management platform like Gremlin, they help you proactively meet your SLOs so you can meet your legal obligations and deliver the perfect customer experience.

View Video

Gremlin

Read more about Achieving SLO Success with Golden Signals and Reliability Testing

5 essential resilience tests for a successful cloud migration

Aug 8, 2024 By Gremlin In Gremlin

Part of the Gremlin Office Hours series: A monthly deep dive with Gremlin experts. Migrating to the cloud usually means faster deployments and easier scalability, but it also means latency. Cloud applications communicate over distributed networks, and while these networks are fast, little bits of latency can quickly add up.

View Video