# Disaster Recovery

Disaster recovery is the planning and capability to restore critical data, information systems, and technology infrastructure after a disruptive event.

This is one page of public article previews, not the complete archive. Follow Next page to continue. Summaries are not the original full articles.

## Kubernetes Backup and Disaster Recovery: What to Protect and How?

DevFeed: [Kubernetes Backup and Disaster Recovery: What to Protect and How?](<https://devfeed.tech/articles/kubernetes-backup-and-disaster-recovery-what-to-protect-and-how-17647.md>)

Original publisher: [Read original article](<https://www.urolime.com/blogs/kubernetes-backup-and-disaster-recovery-what-to-protect-and-how/>)

Author: Urolime Technologies

Published: 2026-09-04T14:50:25Z

Content type: article

Language: en

Sources: [Kubernetes Archives - Urolime Blogs](<https://devfeed.tech/sources/kubernetes-archives-urolime-blogs.md>)

Topics: [Kubernetes](<https://devfeed.tech/topics/kubernetes.md>), [Disaster Recovery](<https://devfeed.tech/topics/disaster-recovery.md>), [Availability](<https://devfeed.tech/topics/availability.md>), [Cloud Native Ecosystem](<https://devfeed.tech/topics/cloud-native-ecosystem.md>)

Tags: [applications](<https://devfeed.tech/tags/applications.md>), [availability](<https://devfeed.tech/tags/availability.md>), [backup](<https://devfeed.tech/tags/backup.md>), [cloud-native](<https://devfeed.tech/tags/cloud-native.md>), [disaster-recovery](<https://devfeed.tech/tags/disaster-recovery.md>), [kubernetes](<https://devfeed.tech/tags/kubernetes.md>), [kubernetes-backup-and-disaster-recovery](<https://devfeed.tech/tags/kubernetes-backup-and-disaster-recovery.md>)

### AI overview

The article explains that Kubernetes high availability and self-healing do not provide disaster recovery. It describes why Kubernetes backup is complex and identifies persistent data, Kubernetes resources, configurations, secrets, and cluster configuration as protection targets, along with several disaster scenarios.

### Source excerpt

Kubernetes has been the bedrock on which most modern-day cloud-native apps have been built- thus providing organizations with a highly efficient way of deploying and managing their workloads. The self-healing nature, automatic scaling, and high availability features of Kubernetes give a false sense that applications are inherently resilient against failures. However, high availability is not [...]

## Push-button migration from Confluent to Redpanda with Shadowing

DevFeed: [Push-button migration from Confluent to Redpanda with Shadowing](<https://devfeed.tech/articles/push-button-migration-from-confluent-to-redpanda-with-shadowing-12715.md>)

Original publisher: [Read original article](<https://www.redpanda.com/blog/migrate-confluent-redpanda-shadowing>)

Author: Trevor Blackford

Published: 2026-08-25T00:00:00Z

Content type: article

Language: en

Sources: [Redpanda](<https://devfeed.tech/sources/redpanda.md>)

Topics: [migration](<https://devfeed.tech/topics/migration.md>), [Replication](<https://devfeed.tech/topics/replication.md>), [Kafka](<https://devfeed.tech/topics/kafka.md>), [Confluent Cloud](<https://devfeed.tech/topics/confluent-cloud.md>), [Confluent Platform](<https://devfeed.tech/topics/confluent-platform.md>), [Disaster Recovery](<https://devfeed.tech/topics/disaster-recovery.md>), [Usability](<https://devfeed.tech/topics/usability.md>)

Tags: [big-bang](<https://devfeed.tech/tags/big-bang.md>), [confluent-cloud](<https://devfeed.tech/tags/confluent-cloud.md>), [confluent-platform](<https://devfeed.tech/tags/confluent-platform.md>), [data-centers](<https://devfeed.tech/tags/data-centers.md>), [disaster-recovery](<https://devfeed.tech/tags/disaster-recovery.md>), [kafka](<https://devfeed.tech/tags/kafka.md>), [migration](<https://devfeed.tech/tags/migration.md>), [replication](<https://devfeed.tech/tags/replication.md>), [schema](<https://devfeed.tech/tags/schema.md>)

### AI overview

The article presents Redpanda Shadowing 26.2 as a low-risk migration path from Confluent Cloud, Confluent Platform, or other Apache Kafka-compatible clusters. It replicates topic data, schemas, offsets, and ACLs while preserving offsets, allowing teams to validate workloads and move applications individually instead of using a big-bang cutover.

### Source excerpt

Migrate off Confluent without the "big cutover weekend." Redpanda Shadowing carries topic data, schemas, offsets, and ACLs on a single link. Available on Self-Managed , BYOC, and Dedicated.

## Multi-Region PostgreSQL Disaster Recovery and Failback with Crunchy PGO

DevFeed: [Multi-Region PostgreSQL Disaster Recovery and Failback with Crunchy PGO](<https://devfeed.tech/articles/multi-region-postgresql-disaster-recovery-and-failback-with-crunchy-pgo-14489.md>)

Original publisher: [Read original article](<https://www.cybertec-postgresql.com/en/multi-region-postgresql-disaster-recovery-and-failback-with-crunchy-pgo/>)

Author: Wellingtone Luvonga

Published: 2026-08-11T05:00:00Z

Content type: tutorial

Language: en

Sources: [CYBERTEC PostgreSQL | Services & Support](<https://devfeed.tech/sources/cybertec-postgresql-services-support.md>)

Topics: [PostgreSQL](<https://devfeed.tech/topics/postgresql.md>), [Disaster Recovery](<https://devfeed.tech/topics/disaster-recovery.md>), [Kubernetes](<https://devfeed.tech/topics/kubernetes.md>), [MinIO](<https://devfeed.tech/topics/minio.md>), [nginx](<https://devfeed.tech/topics/nginx.md>), [TLS (Transport Layer Security)](<https://devfeed.tech/topics/tls.md>), [Amazon S3](<https://devfeed.tech/topics/amazon-s3.md>), [Availability](<https://devfeed.tech/topics/availability.md>)

Tags: [availability](<https://devfeed.tech/tags/availability.md>), [backup](<https://devfeed.tech/tags/backup.md>), [data-recovery](<https://devfeed.tech/tags/data-recovery.md>), [disaster-recovery](<https://devfeed.tech/tags/disaster-recovery.md>), [failover](<https://devfeed.tech/tags/failover.md>), [high-availability](<https://devfeed.tech/tags/high-availability.md>), [how-to](<https://devfeed.tech/tags/how-to.md>), [kubernetes](<https://devfeed.tech/tags/kubernetes.md>), [minio](<https://devfeed.tech/tags/minio.md>), [nginx](<https://devfeed.tech/tags/nginx.md>), [postgresql](<https://devfeed.tech/tags/postgresql.md>), [recovery](<https://devfeed.tech/tags/recovery.md>), [s3](<https://devfeed.tech/tags/s3.md>), [tls](<https://devfeed.tech/tags/tls.md>), [tutorials](<https://devfeed.tech/tags/tutorials.md>)

### AI overview

A tutorial on using Crunchy PostgreSQL Operator on Kubernetes to implement multi-region PostgreSQL disaster recovery and failback. It covers secure MinIO-backed WAL storage through an NGINX TLS reverse proxy, regional failover, and reversing roles without timeline conflicts or archive poisoning.

### Source excerpt

This blog talks about multiregion PostgreSQL disaster recovery using Crunchy PGO. It has some detailed steps to understand better. The post Multi-Region PostgreSQL Disaster Recovery and Failback with Crunchy PGO appeared first on CYBERTEC PostgreSQL | Services & Support.

## Disaster Recovery Testing Best Practices for 2026

DevFeed: [Disaster Recovery Testing Best Practices for 2026](<https://devfeed.tech/articles/disaster-recovery-testing-best-practices-for-2026-13393.md>)

Original publisher: [Read original article](<https://www.harness.io/blog/disaster-recovery-testing-best-practices-how-to-build-a-metrics-driven-resilience-program-in-2026>)

Author: Pritesh Kiri

Published: 2026-08-04T00:00:00Z

Content type: article

Language: en

Sources: [Harness Blog](<https://devfeed.tech/sources/harness-blog.md>)

Topics: [Disaster Recovery](<https://devfeed.tech/topics/disaster-recovery.md>), [Testing](<https://devfeed.tech/topics/testing.md>), [Resilience](<https://devfeed.tech/topics/resilience.md>), [Automation](<https://devfeed.tech/topics/automation.md>)

Tags: [2026](<https://devfeed.tech/tags/2026.md>), [audit](<https://devfeed.tech/tags/audit.md>), [automated](<https://devfeed.tech/tags/automated.md>), [automation](<https://devfeed.tech/tags/automation.md>), [best-practices](<https://devfeed.tech/tags/best-practices.md>), [cross-functional-teams](<https://devfeed.tech/tags/cross-functional-teams.md>), [disaster-recovery](<https://devfeed.tech/tags/disaster-recovery.md>), [external](<https://devfeed.tech/tags/external.md>), [metrics](<https://devfeed.tech/tags/metrics.md>), [pipelines](<https://devfeed.tech/tags/pipelines.md>), [recovery](<https://devfeed.tech/tags/recovery.md>), [resilience](<https://devfeed.tech/tags/resilience.md>), [testing](<https://devfeed.tech/tags/testing.md>)

### AI overview

This article explains how to evolve disaster recovery testing from reactive exercises into a mature, metrics-driven resilience program. It covers risk-aligned testing schedules, automation, continuous improvement, and metrics for demonstrating recovery effectiveness.

### Source excerpt

Learn how to move from ad hoc DR tests to a mature, metrics-driven resilience program with risk-aligned schedules, automation, and the KPIs that prove your reco | Blog

## Announcing Upvest as Confluent's 2026 EMEA Data Streaming Startup of the Year

DevFeed: [Announcing Upvest as Confluent's 2026 EMEA Data Streaming Startup of the Year](<https://devfeed.tech/articles/announcing-upvest-as-confluent-s-2026-emea-data-streaming-startup-of-the-year-11548.md>)

Original publisher: [Read original article](<https://www.confluent.io/blog/announcing-upvest-as-confluents-2026-emea-data-streaming-startup-of-the-year/>)

Author: Tim Graczewski

Published: 2026-07-16T00:07:01Z

Content type: article

Language: en

Sources: [Confluent: Data in motion](<https://devfeed.tech/sources/confluent-data-in-motion.md>)

Topics: [Streaming](<https://devfeed.tech/topics/streaming.md>), [real-time](<https://devfeed.tech/topics/real-time.md>), [Resilience](<https://devfeed.tech/topics/resilience.md>), [event driven](<https://devfeed.tech/topics/event-driven.md>), [Scalability](<https://devfeed.tech/topics/scalability.md>), [Cloud](<https://devfeed.tech/topics/cloud.md>), [data-governance](<https://devfeed.tech/topics/data-governance.md>), [Microservices](<https://devfeed.tech/topics/microservices.md>), [Architecture & Design](<https://devfeed.tech/topics/architecture-design.md>), [Disaster Recovery](<https://devfeed.tech/topics/disaster-recovery.md>)

Tags: [2026](<https://devfeed.tech/tags/2026.md>), [cloud](<https://devfeed.tech/tags/cloud.md>), [cloud-native](<https://devfeed.tech/tags/cloud-native.md>), [compliance](<https://devfeed.tech/tags/compliance.md>), [confluent](<https://devfeed.tech/tags/confluent.md>), [confluent-cloud](<https://devfeed.tech/tags/confluent-cloud.md>), [data](<https://devfeed.tech/tags/data.md>), [data-governance](<https://devfeed.tech/tags/data-governance.md>), [disaster-recovery](<https://devfeed.tech/tags/disaster-recovery.md>), [emea](<https://devfeed.tech/tags/emea.md>), [event-driven](<https://devfeed.tech/tags/event-driven.md>), [fintech](<https://devfeed.tech/tags/fintech.md>), [infrastructure](<https://devfeed.tech/tags/infrastructure.md>), [microservices](<https://devfeed.tech/tags/microservices.md>), [real-time](<https://devfeed.tech/tags/real-time.md>), [resilience](<https://devfeed.tech/tags/resilience.md>), [revolut](<https://devfeed.tech/tags/revolut.md>), [scalability](<https://devfeed.tech/tags/scalability.md>), [startup](<https://devfeed.tech/tags/startup.md>), [streaming](<https://devfeed.tech/tags/streaming.md>)

### AI overview

Confluent recognizes Berlin-based Upvest as its inaugural EMEA Data Streaming Startup of the Year. The article describes Upvest's real-time, event-driven fintech infrastructure, which supports embedded investments, client onboarding, data governance, disaster recovery, scalability, and regulatory resilience for major financial institutions.

### Source excerpt

Confluent for Startups provides an easy on-ramp to Confluent Cloud for early stage startups with great data streaming use cases.

## Top System Design Performance Metrics

DevFeed: [Top System Design Performance Metrics](<https://devfeed.tech/articles/top-system-design-performance-metrics-34692.md>)

Original publisher: [Read original article](<https://newsletter.systemdesigncodex.com/p/top-system-design-performance-metrics>)

Author: Saurabh Dashora

Published: 2026-07-14T08:36:41Z

Content type: tutorial

Language: en

Sources: [System Design Codex](<https://devfeed.tech/sources/system-design-codex.md>)

Topics: [benchmarking](<https://devfeed.tech/topics/benchmarking.md>), [Availability](<https://devfeed.tech/topics/availability.md>), [Architecture & Design](<https://devfeed.tech/topics/architecture-design.md>), [Scalability](<https://devfeed.tech/topics/scalability.md>), [Load Balancing](<https://devfeed.tech/topics/load-balancing.md>), [health checks](<https://devfeed.tech/topics/health-checks.md>), [Concurrency](<https://devfeed.tech/topics/concurrency.md>), [Disaster Recovery](<https://devfeed.tech/topics/disaster-recovery.md>), [Database](<https://devfeed.tech/topics/database.md>), [sharding](<https://devfeed.tech/topics/sharding.md>), [IO](<https://devfeed.tech/topics/io.md>), [API](<https://devfeed.tech/topics/api.md>)

Tags: [api](<https://devfeed.tech/tags/api.md>), [availability](<https://devfeed.tech/tags/availability.md>), [blocking](<https://devfeed.tech/tags/blocking.md>), [concurrency](<https://devfeed.tech/tags/concurrency.md>), [database](<https://devfeed.tech/tags/database.md>), [disaster-recovery](<https://devfeed.tech/tags/disaster-recovery.md>), [health-checks](<https://devfeed.tech/tags/health-checks.md>), [load-balancing](<https://devfeed.tech/tags/load-balancing.md>), [performance](<https://devfeed.tech/tags/performance.md>), [performance-metrics](<https://devfeed.tech/tags/performance-metrics.md>), [scalability](<https://devfeed.tech/tags/scalability.md>), [sharding](<https://devfeed.tech/tags/sharding.md>), [system-design](<https://devfeed.tech/tags/system-design.md>)

### AI overview

A tutorial on system design performance metrics, focusing on availability and throughput. It explains how these metrics are measured and outlines techniques such as load balancing, health checks, failover, redundancy, disaster recovery, query optimization, sharding, and asynchronous processing.

### Source excerpt

Must Know Metrics

## How to Automate CockroachDB Operations with AI Using Cockroach University

DevFeed: [How to Automate CockroachDB Operations with AI Using Cockroach University](<https://devfeed.tech/articles/how-to-automate-cockroachdb-operations-with-ai-using-cockroach-university-23744.md>)

Original publisher: [Read original article](<https://cockroachlabs.com/blog/ai-cockroachdb-training-courses>)

Author: Nathan Zamecnik

Published: 2026-07-13T00:00:00Z

Content type: article

Language: en

Sources: [Cockroach Labs](<https://devfeed.tech/sources/cockroach-labs.md>)

Topics: [CockroachDB](<https://devfeed.tech/topics/cockroachdb.md>), [Agent Skills](<https://devfeed.tech/topics/agent-skills.md>), [Artificial Intelligence](<https://devfeed.tech/topics/ai.md>), [maintenance](<https://devfeed.tech/topics/maintenance.md>), [Claude Code](<https://devfeed.tech/topics/claude-code.md>), [Disaster Recovery](<https://devfeed.tech/topics/disaster-recovery.md>), [Security](<https://devfeed.tech/topics/security.md>), [observability](<https://devfeed.tech/topics/observability.md>)

Tags: [agent-skills](<https://devfeed.tech/tags/agent-skills.md>), [agents](<https://devfeed.tech/tags/agents.md>), [ai-agents](<https://devfeed.tech/tags/ai-agents.md>), [claude-code](<https://devfeed.tech/tags/claude-code.md>), [cockroachdb](<https://devfeed.tech/tags/cockroachdb.md>), [disaster-recovery](<https://devfeed.tech/tags/disaster-recovery.md>), [maintenance](<https://devfeed.tech/tags/maintenance.md>), [observability](<https://devfeed.tech/tags/observability.md>), [security](<https://devfeed.tech/tags/security.md>)

### AI overview

The article describes CockroachDB Agent Skills, a public repository of machine-executable operational workflows that AI agents such as Claude Code and Cursor can use. It also introduces three Cockroach University courses covering rolling restarts, encryption at rest, and planned node maintenance, along with a free trial of the CockroachDB Training Subscription.

### Source excerpt

AI agents aren't just changing how applications interact with databases, they're changing how teams operate them.

## Me and my shadow (link!): Disaster recovery replication made easy

DevFeed: [Me and my shadow (link!): Disaster recovery replication made easy](<https://devfeed.tech/articles/me-and-my-shadow-link-disaster-recovery-replication-made-easy-12770.md>)

Original publisher: [Read original article](<https://www.redpanda.com/blog/shadow-linking-disaster-recovery-replication-made-easy>)

Author: Paul Wilkinson

Published: 2026-04-21T00:00:00Z

Content type: article

Language: en

Sources: [Redpanda](<https://devfeed.tech/sources/redpanda.md>)

Topics: [Disaster Recovery](<https://devfeed.tech/topics/disaster-recovery.md>), [Replication](<https://devfeed.tech/topics/replication.md>), [Streaming](<https://devfeed.tech/topics/streaming.md>), [systems](<https://devfeed.tech/topics/systems.md>), [API](<https://devfeed.tech/topics/api.md>), [Kafka](<https://devfeed.tech/topics/kafka.md>)

Tags: [disaster-recovery](<https://devfeed.tech/tags/disaster-recovery.md>), [enterprise](<https://devfeed.tech/tags/enterprise.md>), [failover](<https://devfeed.tech/tags/failover.md>), [kafka](<https://devfeed.tech/tags/kafka.md>), [operational](<https://devfeed.tech/tags/operational.md>), [outage](<https://devfeed.tech/tags/outage.md>), [real-time](<https://devfeed.tech/tags/real-time.md>), [replication](<https://devfeed.tech/tags/replication.md>), [streaming](<https://devfeed.tech/tags/streaming.md>)

### AI overview

This article explains Redpanda Shadow Linking as a built-in disaster recovery feature for continuous replication between streaming clusters. It describes an active-passive architecture in which a read-only shadow cluster receives topics, schemas, offsets, ACLs, commits, and messages, then becomes writable after failover.

### Source excerpt

Shadow Linking made simple. Get started with real-time replication.

## Introducing the SQL Server on Kubernetes Operator

DevFeed: [Introducing the SQL Server on Kubernetes Operator](<https://devfeed.tech/articles/introducing-the-sql-server-on-kubernetes-operator-17552.md>)

Original publisher: [Read original article](<https://www.nocentino.com/posts/2026-04-12-introducing-sql-on-k8s-operator/>)

Author: Anthony Nocentino

Published: 2026-04-12T05:00:00Z

Content type: tutorial

Language: en

Sources: [Kubernetes on Anthony Nocentino's Blog](<https://devfeed.tech/sources/kubernetes-on-anthony-nocentino-s-blog.md>)

Topics: [Kubernetes](<https://devfeed.tech/topics/kubernetes.md>), [sql-server](<https://devfeed.tech/topics/sql-server.md>), [Availability](<https://devfeed.tech/topics/availability.md>), [Disaster Recovery](<https://devfeed.tech/topics/disaster-recovery.md>), [Open Source](<https://devfeed.tech/topics/open-source.md>)

Tags: [automated](<https://devfeed.tech/tags/automated.md>), [automatic](<https://devfeed.tech/tags/automatic.md>), [availability](<https://devfeed.tech/tags/availability.md>), [availability-groups](<https://devfeed.tech/tags/availability-groups.md>), [containers](<https://devfeed.tech/tags/containers.md>), [disaster-recovery](<https://devfeed.tech/tags/disaster-recovery.md>), [failover](<https://devfeed.tech/tags/failover.md>), [high-availability](<https://devfeed.tech/tags/high-availability.md>), [kubernetes](<https://devfeed.tech/tags/kubernetes.md>), [open-source](<https://devfeed.tech/tags/open-source.md>), [operator](<https://devfeed.tech/tags/operator.md>), [sql-server](<https://devfeed.tech/tags/sql-server.md>)

### AI overview

This tutorial introduces sql-on-k8s-operator, an open-source Kubernetes operator for managing SQL Server workloads. It addresses high availability and disaster recovery by automating lifecycle management, failover, and Always On Availability Group setup, and demonstrates deployment with custom resources.

### Source excerpt

Are you considering replatforming your SQL Server workload due to recent vendor changes, but still need high availability and disaster recovery? You're not alone. One of the challenges with running SQL Server on Kubernetes is that there's no Kubernetes operator available. That means no automated lifecycle management, no automatic failover, and no standard way to bootstrap an Always On Availability Group on Kubernetes. I'm excited to share it today as an open-source project: sql-on-k8s-operator. Let's go.

## v18.2.8 Reef released

DevFeed: [v18.2.8 Reef released](<https://devfeed.tech/articles/v18-2-8-reef-released-12339.md>)

Original publisher: [Read original article](<https://ceph.io/en/news/blog/2026/v18-2-8-reef-released/>)

Author: Yuri Weinstein

Published: 2026-03-20T00:00:00Z

Content type: release

Language: en

Sources: [Ceph Blog](<https://devfeed.tech/sources/ceph-blog.md>)

Topics: [Vulnerabilities](<https://devfeed.tech/topics/vulnerabilities.md>), [Security](<https://devfeed.tech/topics/security.md>), [configuration](<https://devfeed.tech/topics/configuration.md>), [SSL](<https://devfeed.tech/topics/ssl.md>), [Disaster Recovery](<https://devfeed.tech/topics/disaster-recovery.md>), [data](<https://devfeed.tech/topics/data.md>)

Tags: [2026](<https://devfeed.tech/tags/2026.md>), [aws](<https://devfeed.tech/tags/aws.md>), [blog-post](<https://devfeed.tech/tags/blog-post.md>), [bug](<https://devfeed.tech/tags/bug.md>), [configuration](<https://devfeed.tech/tags/configuration.md>), [disaster-recovery](<https://devfeed.tech/tags/disaster-recovery.md>), [en-article](<https://devfeed.tech/tags/en-article.md>), [en-blog-post](<https://devfeed.tech/tags/en-blog-post.md>), [reef](<https://devfeed.tech/tags/reef.md>), [release](<https://devfeed.tech/tags/release.md>), [security](<https://devfeed.tech/tags/security.md>), [ssl](<https://devfeed.tech/tags/ssl.md>), [vulnerability](<https://devfeed.tech/tags/vulnerability.md>)

### AI overview

Ceph v18.2.8 Reef is the eighth and expected final backport release in the Reef series, released on March 20, 2026. It includes upgrade guidance, security fixes for CephFS clients and Mgr Alerts, RGW and CephFS reliability fixes, configurable JWKS URL verification for AWS compliance, and additional MDS and CephFS changes.

### Source excerpt

This is the eighth, and expected to be last, backport release in the Reef series. We recommend that all users update to this release. Release Date ¶ March 20, 2026 Known Issues ¶ During QA for v18.2.8, it was found that there was a bug for upgrades from Pacific to Reef. Pacific OSDs (and other Ceph daemons) were still using a deprecated connection feature bit that was adopted to indicate a Reef OSD. This can cause a OSD_UPGRADE_FINISHED warning before all OSDs are actually upgraded to Reef. There are no known issues associated with Pacific and Reef OSDs interoperating where Pacific OSDs are "advertising" Reef compatibility; however, out of an abundance of caution, we no longer recommend upgrading from Pacific to Reef directly. Security Fixes ¶ CephFS Client: A fix was merged to prohibit unprivileged users from modifying the sgid or suid bits on a file. Previously, unprivileged users were inadvertently permitted to set these bits if they were the sole bits being modified. Mgr Alerts: The SMTP SSL context was enforced in the mgr/alerts module to resolve a security vulnerability (GHSA-xj9f-7g59-m4jx). Notable Changes ¶ RGW (RADOS Gateway): Fixed an issue where bucket rm --bypass-gc was mistakenly removing head objects instead of tail objects, potentially causing data inconsistencies. Fixed rgw-restore-bucket-index to handle objects with leading hyphens and to process versioned buckets correctly. Addressed an issue in the msg/async protocol that caused memory locks and hangs during connection shutdown. RGW STS: Made JWKS URL verification configurable for AWS compliance via the rgw_enable_jwks_url_verification configuration. CephFS / MDS: Prevented the MDS from stalling (up to 5 seconds) during rename/stat workloads by forcing the log to nudge for unstable locks after early replies. Fixed cephfs-journal-tool so it no longer incorrectly resets the journal trim position during disaster recovery, which was causing stale journal objects to linger forever in the metadata pool

## How Gremlin makes disaster recovery testing easier and faster

DevFeed: [How Gremlin makes disaster recovery testing easier and faster](<https://devfeed.tech/articles/how-gremlin-makes-disaster-recovery-testing-easier-and-faster-11593.md>)

Original publisher: [Read original article](<https://www.gremlin.com/blog/how-gremlin-makes-disaster-recovery-testing-easier-and-faster>)

Author: Gavin Cahill

Published: 2026-03-04T00:00:00Z

Content type: tutorial

Language: en

Sources: [Gremlin Blog](<https://devfeed.tech/sources/gremlin-blog.md>)

Topics: [Disaster Recovery](<https://devfeed.tech/topics/disaster-recovery.md>), [Resilience](<https://devfeed.tech/topics/resilience.md>), [Testing](<https://devfeed.tech/topics/testing.md>), [Cloud](<https://devfeed.tech/topics/cloud.md>)

Tags: [backup](<https://devfeed.tech/tags/backup.md>), [cloud](<https://devfeed.tech/tags/cloud.md>), [disaster-recovery](<https://devfeed.tech/tags/disaster-recovery.md>), [failover](<https://devfeed.tech/tags/failover.md>), [gremlin](<https://devfeed.tech/tags/gremlin.md>), [resilience](<https://devfeed.tech/tags/resilience.md>), [testing](<https://devfeed.tech/tags/testing.md>), [use-cases](<https://devfeed.tech/tags/use-cases.md>)

### AI overview

The article explains how Gremlin's Disaster Recovery Testing helps teams test disaster recovery plans by simulating failures such as zone evacuations, region failovers, and cloud-provider outages. It recommends establishing service baselines with test suites, running tests regularly, and repeating them to verify fixes.

### Source excerpt

Gremlin's Disaster Recovery Testing makes it easy to run zone evacuations, region failovers, and more for a fraction of the lift of traditional disaster recovery testing.

## Aloha, Redpanda FY27 company kick-off

DevFeed: [Aloha, Redpanda FY27 company kick-off](<https://devfeed.tech/articles/aloha-redpanda-fy27-company-kick-off-12744.md>)

Original publisher: [Read original article](<https://www.redpanda.com/blog/redpanda-company-kickoff-fy27>)

Author: Jenny Medeiros

Published: 2026-02-24T00:00:00Z

Content type: article

Language: en

Sources: [Redpanda](<https://devfeed.tech/sources/redpanda.md>)

Topics: [AI Infrastructure](<https://devfeed.tech/topics/ai-infrastructure.md>), [Redpanda-Connect](<https://devfeed.tech/topics/redpanda-connect.md>), [Serverless](<https://devfeed.tech/topics/serverless.md>), [Amazon Web Services](<https://devfeed.tech/topics/aws.md>), [Disaster Recovery](<https://devfeed.tech/topics/disaster-recovery.md>), [MCP Server](<https://devfeed.tech/topics/mcp-server.md>), [Graphs](<https://devfeed.tech/topics/graphs.md>), [Google](<https://devfeed.tech/topics/google.md>)

Tags: [ai](<https://devfeed.tech/tags/ai.md>), [ai-infrastructure](<https://devfeed.tech/tags/ai-infrastructure.md>), [aws](<https://devfeed.tech/tags/aws.md>), [disaster-recovery](<https://devfeed.tech/tags/disaster-recovery.md>), [google](<https://devfeed.tech/tags/google.md>), [graphs](<https://devfeed.tech/tags/graphs.md>), [mcp-server](<https://devfeed.tech/tags/mcp-server.md>), [redpanda-connect](<https://devfeed.tech/tags/redpanda-connect.md>), [serverless](<https://devfeed.tech/tags/serverless.md>)

### AI overview

Redpanda's FY27 company kick-off brought more than 160 employees together in San Diego to reflect on the company's progress and direction. The article highlights the Agentic Data Plane, Redpanda Cloud, Redpanda Serverless on AWS, disaster recovery in Redpanda 25.3, MCP server development, and new Redpanda Connect connectors, alongside talks about company culture and using AI to reshape work.

### Source excerpt

From AI hackathons to hakas, here's what happened at our fourth company kick-off.

## Temporal raises $300M Series D at a $5B valuation as AI drives demand for Durable Execution

DevFeed: [Temporal raises $300M Series D at a $5B valuation as AI drives demand for Durable Execution](<https://devfeed.tech/articles/temporal-raises-300m-series-d-at-a-5b-valuation-as-ai-drives-demand-for-durable-execution-36025.md>)

Original publisher: [Read original article](<https://temporal.io/blog/temporal-raises-usd300m-series-d-at-a-usd5b-valuation>)

Author: Allanah Hughes

Published: 2026-02-17T00:00:00Z

Content type: release

Language: en

Sources: [Temporal Blog](<https://devfeed.tech/sources/temporal-blog.md>)

Topics: [distributed-systems](<https://devfeed.tech/topics/distributed-systems.md>), [Artificial Intelligence](<https://devfeed.tech/topics/ai.md>), [Orchestration](<https://devfeed.tech/topics/orchestration.md>), [Cloud](<https://devfeed.tech/topics/cloud.md>), [Disaster Recovery](<https://devfeed.tech/topics/disaster-recovery.md>), [backends](<https://devfeed.tech/topics/backends.md>)

Tags: [agentic-ai](<https://devfeed.tech/tags/agentic-ai.md>), [ai](<https://devfeed.tech/tags/ai.md>), [ai-agents](<https://devfeed.tech/tags/ai-agents.md>), [announcements](<https://devfeed.tech/tags/announcements.md>), [backends](<https://devfeed.tech/tags/backends.md>), [cloud](<https://devfeed.tech/tags/cloud.md>), [disaster-recovery](<https://devfeed.tech/tags/disaster-recovery.md>), [distributed-systems](<https://devfeed.tech/tags/distributed-systems.md>), [fault-tolerance](<https://devfeed.tech/tags/fault-tolerance.md>), [orchestration](<https://devfeed.tech/tags/orchestration.md>)

### AI overview

Temporal announces a $300 million Series D at a $5 billion post-money valuation, led by Andreessen Horowitz. The company says demand for durable execution is increasing as AI and other long-running production workflows require state preservation and recoverable failures.

### Source excerpt

Temporal raises $300M Series D at a $5B valuation as AI drives demand for Durable Execution.

## Announcing Disaster Recovery Testing

DevFeed: [Announcing Disaster Recovery Testing](<https://devfeed.tech/articles/announcing-disaster-recovery-testing-11558.md>)

Original publisher: [Read original article](<https://www.gremlin.com/blog/announcing-disaster-recovery-testing>)

Author: Andre Newman

Published: 2026-02-03T00:00:00Z

Content type: release

Language: en

Sources: [Gremlin Blog](<https://devfeed.tech/sources/gremlin-blog.md>)

Topics: [Disaster Recovery](<https://devfeed.tech/topics/disaster-recovery.md>), [Testing](<https://devfeed.tech/topics/testing.md>), [Incident response](<https://devfeed.tech/topics/incident-response.md>), [Resilience](<https://devfeed.tech/topics/resilience.md>), [Availability](<https://devfeed.tech/topics/availability.md>), [Cloud](<https://devfeed.tech/topics/cloud.md>), [datacenter](<https://devfeed.tech/topics/datacenter.md>), [Amazon Web Services](<https://devfeed.tech/topics/aws.md>), [Azure](<https://devfeed.tech/topics/azure.md>)

Tags: [announcements](<https://devfeed.tech/tags/announcements.md>), [availability](<https://devfeed.tech/tags/availability.md>), [aws](<https://devfeed.tech/tags/aws.md>), [azure](<https://devfeed.tech/tags/azure.md>), [disaster-recovery](<https://devfeed.tech/tags/disaster-recovery.md>), [failover](<https://devfeed.tech/tags/failover.md>), [gremlin](<https://devfeed.tech/tags/gremlin.md>), [health-checks](<https://devfeed.tech/tags/health-checks.md>), [incident-response](<https://devfeed.tech/tags/incident-response.md>), [infrastructure](<https://devfeed.tech/tags/infrastructure.md>), [metrics](<https://devfeed.tech/tags/metrics.md>), [resilience](<https://devfeed.tech/tags/resilience.md>), [systems](<https://devfeed.tech/tags/systems.md>), [testing](<https://devfeed.tech/tags/testing.md>)

### AI overview

Gremlin announces Disaster Recovery Testing, a feature for running organization-wide zone, region, and datacenter-scale experiments. It helps teams validate failover, disaster recovery, and incident response processes, with health checks that can automatically halt tests when key metrics exceed defined SLA limits.

### Source excerpt

Gremlin announces Disaster Recovery Testing for validating region failover processes, disaster recovery plans, incident response procedures, and more.

## 20 years of Getting Things Done

DevFeed: [20 years of Getting Things Done](<https://devfeed.tech/articles/20-years-of-getting-things-done-27730.md>)

Original publisher: [Read original article](<https://gagor.pro/2025/11/20-years-of-getting-things-done/>)

Author: Tom

Published: 2025-11-27T00:00:00Z

Content type: article

Language: en

Sources: [Tomasz Gągor](<https://devfeed.tech/sources/tomasz-gagor.md>)

Topics: [systems](<https://devfeed.tech/topics/systems.md>), [DevOps](<https://devfeed.tech/topics/devops.md>), [Monitoring](<https://devfeed.tech/topics/monitoring.md>), [backups](<https://devfeed.tech/topics/backups.md>), [health checks](<https://devfeed.tech/topics/health-checks.md>), [Disaster Recovery](<https://devfeed.tech/topics/disaster-recovery.md>)

Tags: [2](<https://devfeed.tech/tags/2.md>), [advice](<https://devfeed.tech/tags/advice.md>), [backup](<https://devfeed.tech/tags/backup.md>), [best-practices](<https://devfeed.tech/tags/best-practices.md>), [devops](<https://devfeed.tech/tags/devops.md>), [getting-things-done](<https://devfeed.tech/tags/getting-things-done.md>), [gtd](<https://devfeed.tech/tags/gtd.md>), [leadership](<https://devfeed.tech/tags/leadership.md>), [learnings](<https://devfeed.tech/tags/learnings.md>), [linux](<https://devfeed.tech/tags/linux.md>), [monitoring](<https://devfeed.tech/tags/monitoring.md>), [productivity](<https://devfeed.tech/tags/productivity.md>), [recovery](<https://devfeed.tech/tags/recovery.md>), [time-management](<https://devfeed.tech/tags/time-management.md>), [workshop](<https://devfeed.tech/tags/workshop.md>)

### AI overview

A reflection on 20 years of using Getting Things Done, including how external systems shaped the author's work as a system administrator and DevOps engineer. It describes replacing manual backup checks with failure notifications and introducing monitoring and health checks to move operations from reactive to proactive.

### Source excerpt

Reflections and practical tips from 20 years using Getting Things Done, plus a short workshop plan for leaders who want actionable routines.

## Durable under pressure: How developers kept running during the AWS us-east-1 outage

DevFeed: [Durable under pressure: How developers kept running during the AWS us-east-1 outage](<https://devfeed.tech/articles/durable-under-pressure-how-developers-kept-running-during-the-aws-us-east-1-outage-35851.md>)

Original publisher: [Read original article](<https://temporal.io/blog/how-devs-kept-running-during-the-aws-us-east-1-oct-20-2025>)

Author: Luke Knepper

Published: 2025-11-07T00:00:00Z

Content type: article

Language: en

Sources: [Temporal Blog](<https://devfeed.tech/sources/temporal-blog.md>)

Topics: [Amazon Web Services](<https://devfeed.tech/topics/aws.md>), [Availability](<https://devfeed.tech/topics/availability.md>), [Cloud](<https://devfeed.tech/topics/cloud.md>), [Disaster Recovery](<https://devfeed.tech/topics/disaster-recovery.md>), [Replication](<https://devfeed.tech/topics/replication.md>), [incident](<https://devfeed.tech/topics/incident.md>), [DynamoDB](<https://devfeed.tech/topics/dynamodb.md>)

Tags: [2025](<https://devfeed.tech/tags/2025.md>), [availability](<https://devfeed.tech/tags/availability.md>), [aws](<https://devfeed.tech/tags/aws.md>), [disaster-recovery](<https://devfeed.tech/tags/disaster-recovery.md>), [dynamodb](<https://devfeed.tech/tags/dynamodb.md>), [high-availability](<https://devfeed.tech/tags/high-availability.md>), [incident](<https://devfeed.tech/tags/incident.md>), [outage](<https://devfeed.tech/tags/outage.md>), [temporal-concepts](<https://devfeed.tech/tags/temporal-concepts.md>), [us](<https://devfeed.tech/tags/us.md>)

### AI overview

The article examines the October 20, 2025 AWS us-east-1 outage and explains how multi-region high availability and disaster recovery strategies helped some applications continue running. It describes Temporal Cloud's replication and failover features and FireHydrant's preparation and testing.

### Source excerpt

The AWS us-east-1 outage affected many, but some devs were able to stay afloat. Find out how.

## Redpanda 25.3 delivers near-instant disaster recovery and more

DevFeed: [Redpanda 25.3 delivers near-instant disaster recovery and more](<https://devfeed.tech/articles/redpanda-25-3-delivers-near-instant-disaster-recovery-and-more-12664.md>)

Original publisher: [Read original article](<https://www.redpanda.com/blog/25-3-enterprise-disaster-recovery>)

Author: Matt Schumpert

Published: 2025-11-06T00:00:00Z

Content type: release

Language: en

Sources: [Redpanda](<https://devfeed.tech/sources/redpanda.md>)

Topics: [Disaster Recovery](<https://devfeed.tech/topics/disaster-recovery.md>), [Resilience](<https://devfeed.tech/topics/resilience.md>), [Kafka](<https://devfeed.tech/topics/kafka.md>), [Streaming](<https://devfeed.tech/topics/streaming.md>), [Availability](<https://devfeed.tech/topics/availability.md>), [configuration](<https://devfeed.tech/topics/configuration.md>), [Cloud](<https://devfeed.tech/topics/cloud.md>), [API](<https://devfeed.tech/topics/api.md>), [systems](<https://devfeed.tech/topics/systems.md>)

Tags: [cloud](<https://devfeed.tech/tags/cloud.md>), [clusters](<https://devfeed.tech/tags/clusters.md>), [configuration](<https://devfeed.tech/tags/configuration.md>), [consumer](<https://devfeed.tech/tags/consumer.md>), [disaster-recovery](<https://devfeed.tech/tags/disaster-recovery.md>), [high-availability](<https://devfeed.tech/tags/high-availability.md>), [integration](<https://devfeed.tech/tags/integration.md>), [kafka](<https://devfeed.tech/tags/kafka.md>), [process](<https://devfeed.tech/tags/process.md>), [production](<https://devfeed.tech/tags/production.md>), [real-time](<https://devfeed.tech/tags/real-time.md>), [recovery](<https://devfeed.tech/tags/recovery.md>), [release](<https://devfeed.tech/tags/release.md>), [resilience](<https://devfeed.tech/tags/resilience.md>), [server](<https://devfeed.tech/tags/server.md>), [stream-processing](<https://devfeed.tech/tags/stream-processing.md>), [streaming](<https://devfeed.tech/tags/streaming.md>), [systems](<https://devfeed.tech/tags/systems.md>), [update](<https://devfeed.tech/tags/update.md>)

### AI overview

The article previews Redpanda 25.3, focusing on Shadowing, a disaster recovery and business continuity feature for streaming workloads. Shadowing creates a byte-for-byte, offset-preserving hot standby of a production cluster in another region, including topics, configurations, consumer offsets, ACLs, and schemas. It is designed to provide recovery point and recovery time objectives measured in seconds, with setup through configuration or Redpanda Console and replication using the standard Kafka API.

### Source excerpt

Check out how the latest Redpanda release is providing the resilience, efficiency, and interoperability that modern enterprises need for their streaming workloads.

## DEEP Conference 2025: Digital Twins, Security, and Database Access Presentations

DevFeed: [DEEP Conference 2025: Digital Twins, Security, and Database Access Presentations](<https://devfeed.tech/articles/deep-is-still-a-must-attend-boutique-conference-11256.md>)

Original publisher: [Read original article](<https://blog.ipspace.net/2025/10/deep-conference-2025/>)

Published: 2025-10-24T05:54:00Z

Content type: opinion

Language: en

Sources: [ipSpace.net blog](<https://devfeed.tech/sources/ipspace-net-blog.md>)

Topics: [Disaster Recovery](<https://devfeed.tech/topics/disaster-recovery.md>), [Latency](<https://devfeed.tech/topics/latency.md>), [networking](<https://devfeed.tech/topics/networking.md>), [Testing](<https://devfeed.tech/topics/testing.md>)

Tags: [conference](<https://devfeed.tech/tags/conference.md>), [conferences](<https://devfeed.tech/tags/conferences.md>), [disaster-recovery](<https://devfeed.tech/tags/disaster-recovery.md>), [latency](<https://devfeed.tech/tags/latency.md>), [networking](<https://devfeed.tech/tags/networking.md>), [testing](<https://devfeed.tech/tags/testing.md>)

### AI overview

The author recounts speaking at the 2025 DEEP Conference in Zadar, Croatia, where the presentation covered digital twins for disaster recovery and avoidance testing, including bandwidth and latency emulation. The article also highlights presentations on social engineering and malware, North Korean hacking tools, and controlled access to production databases.

### Source excerpt

I love well-organized small conferences, so it wasn't hard to persuade me to have another talk at the DEEP Conference in Zadar, Croatia. This time, I talked about the role of digital twins in disaster recovery/avoidance testing. You might know my take on networking digital twins; after that, I only had enough time to focus on bandwidth and latency matter, and this is how you emulate limited bandwidth and add latency bit. Read more ...

## Migrating a Platform in 3 Days: The Artisan of the Day Is Dan Ferguson.

DevFeed: [Migrating a Platform in 3 Days: The Artisan of the Day Is Dan Ferguson.](<https://devfeed.tech/articles/migrating-a-platform-in-3-days-the-artisan-of-the-day-is-dan-ferguson-3880.md>)

Original publisher: [Read original article](<https://laravel.com/blog/migrating-a-platform-in-3-days-the-artisan-of-the-day-is-dan-ferguson>)

Author: Ana Tavares

Published: 2025-09-01T14:21:00Z

Content type: article

Language: en

Sources: [Laravel Blog](<https://devfeed.tech/sources/laravel-blog.md>)

Topics: [Laravel](<https://devfeed.tech/topics/laravel.md>), [Caching](<https://devfeed.tech/topics/caching.md>), [Next.js](<https://devfeed.tech/topics/next-js.md>), [React](<https://devfeed.tech/topics/react.md>), [Disaster Recovery](<https://devfeed.tech/topics/disaster-recovery.md>), [Security](<https://devfeed.tech/topics/security.md>), [PHP](<https://devfeed.tech/topics/php.md>)

Tags: [ai](<https://devfeed.tech/tags/ai.md>), [australia](<https://devfeed.tech/tags/australia.md>), [cache](<https://devfeed.tech/tags/cache.md>), [caching](<https://devfeed.tech/tags/caching.md>), [disaster-recovery](<https://devfeed.tech/tags/disaster-recovery.md>), [laravel](<https://devfeed.tech/tags/laravel.md>), [php](<https://devfeed.tech/tags/php.md>), [security](<https://devfeed.tech/tags/security.md>)

### AI overview

Dan Ferguson of Communiti Labs describes migrating a platform from a React and Next.js stack to Laravel in three days. He cites Laravel's built-in support for mail, queues, and caching, along with improved delivery confidence. The article also presents Helping Homes, a Laravel-based disaster-response site that registered more than 5,000 beds during Australia's Black Summer bushfires and was later strengthened with security, testing, caching, and improved page rendering.

### Source excerpt

Dan Ferguson of Communiti Labs uses Laravel to power civic tech, disaster response tools, and AI-driven community analysis platforms.

## What it Takes to Recover a 100 TB Postgres Database in AWS RDS

DevFeed: [What it Takes to Recover a 100 TB Postgres Database in AWS RDS](<https://devfeed.tech/articles/what-it-takes-to-recover-a-100-tb-postgres-database-in-aws-rds-5765.md>)

Original publisher: [Read original article](<https://neon.com/blog/recover-large-postgres-databases>)

Author: Carlota Soto

Published: 2025-02-04T01:22:52Z

Content type: article

Language: en

Sources: [Blog -- Neon Docs](<https://devfeed.tech/sources/blog-neon-docs.md>)

Topics: [Database](<https://devfeed.tech/topics/database.md>), [Amazon Web Services](<https://devfeed.tech/topics/aws.md>), [Disaster Recovery](<https://devfeed.tech/topics/disaster-recovery.md>), [Availability](<https://devfeed.tech/topics/availability.md>), [Replication](<https://devfeed.tech/topics/replication.md>), [Amazon S3](<https://devfeed.tech/topics/amazon-s3.md>)

Tags: [amazon-s3](<https://devfeed.tech/tags/amazon-s3.md>), [availability](<https://devfeed.tech/tags/availability.md>), [aws](<https://devfeed.tech/tags/aws.md>), [backup](<https://devfeed.tech/tags/backup.md>), [database](<https://devfeed.tech/tags/database.md>), [disaster-recovery](<https://devfeed.tech/tags/disaster-recovery.md>), [infrastructure](<https://devfeed.tech/tags/infrastructure.md>), [postgres](<https://devfeed.tech/tags/postgres.md>), [product](<https://devfeed.tech/tags/product.md>), [production](<https://devfeed.tech/tags/production.md>), [replication](<https://devfeed.tech/tags/replication.md>)

### AI overview

This article examines how AWS RDS handles recovery for a 100 TB Postgres database. It distinguishes logical failures, which require point-in-time recovery using snapshots and WAL replay, from infrastructure or Availability Zone failures, which require high availability through redundant instances. It explains that snapshot restoration and WAL replay can take hours at this scale, while Multi-AZ standbys support automatic failover but do not protect against replicated data corruption.

### Source excerpt

Imagine your AWS RDS Postgres database scales to 100 TB. Now, sh*t disaster happens--your production instance experiences a critical failure. How would RDS handle different failure modes? And how long would it take to restore full production capability? Production failure modes Pr...

## Latency Numbers Every Programmer Should Know

DevFeed: [Latency Numbers Every Programmer Should Know](<https://devfeed.tech/articles/latency-numbers-every-programmer-should-know-11101.md>)

Original publisher: [Read original article](<https://blog.ipspace.net/2024/11/worth-reading-latency-numbers-programmers-should-know/>)

Published: 2024-11-08T07:10:00Z

Content type: opinion

Language: en

Sources: [ipSpace.net blog](<https://devfeed.tech/sources/ipspace-net-blog.md>)

Topics: [Latency](<https://devfeed.tech/topics/latency.md>), [networking](<https://devfeed.tech/topics/networking.md>), [Disaster Recovery](<https://devfeed.tech/topics/disaster-recovery.md>), [Website](<https://devfeed.tech/topics/website.md>), [cURL](<https://devfeed.tech/topics/curl.md>)

Tags: [curl](<https://devfeed.tech/tags/curl.md>), [disaster-recovery](<https://devfeed.tech/tags/disaster-recovery.md>), [latency](<https://devfeed.tech/tags/latency.md>), [networking](<https://devfeed.tech/tags/networking.md>), [terminal](<https://devfeed.tech/tags/terminal.md>), [web](<https://devfeed.tech/tags/web.md>), [worth-reading](<https://devfeed.tech/tags/worth-reading.md>)

### AI overview

The author recommends an infographic about latency numbers, highlighting the contrast between SSD read latency and cross-site round-trip time. The article also mentions a website for checking latency from different vantage points and notes that the infographic can be fetched in a terminal with curl.

### Source excerpt

One of the key arguments against stretched clusters (and similar stupidities) I used in my Disaster Recovery Myths presentation was the SSD read latency versus cross-site round-trip time. Thanks to Networking Notes, I found a great infographic I can use in my next presentation (bonus points: it also works great in a terminal when fetched with curl) and a site that checks the latency of your web site from various vantage points.

## How reliability engineering can verify disaster recovery plans

DevFeed: [How reliability engineering can verify disaster recovery plans](<https://devfeed.tech/articles/how-reliability-engineering-can-verify-disaster-recovery-plans-11599.md>)

Original publisher: [Read original article](<https://www.gremlin.com/blog/how-reliability-engineering-can-verify-disaster-recovery-plans>)

Author: Gavin Cahill

Published: 2024-11-05T00:00:00Z

Content type: article

Language: en

Sources: [Gremlin Blog](<https://devfeed.tech/sources/gremlin-blog.md>)

Topics: [Disaster Recovery](<https://devfeed.tech/topics/disaster-recovery.md>), [Resilience](<https://devfeed.tech/topics/resilience.md>), [systems](<https://devfeed.tech/topics/systems.md>), [configuration](<https://devfeed.tech/topics/configuration.md>), [Availability](<https://devfeed.tech/topics/availability.md>), [Networks](<https://devfeed.tech/topics/networks.md>)

Tags: [apra](<https://devfeed.tech/tags/apra.md>), [availability](<https://devfeed.tech/tags/availability.md>), [canada](<https://devfeed.tech/tags/canada.md>), [compliance](<https://devfeed.tech/tags/compliance.md>), [configuration](<https://devfeed.tech/tags/configuration.md>), [cps-230](<https://devfeed.tech/tags/cps-230.md>), [disaster-recovery](<https://devfeed.tech/tags/disaster-recovery.md>), [eu](<https://devfeed.tech/tags/eu.md>), [gremlin](<https://devfeed.tech/tags/gremlin.md>), [outage](<https://devfeed.tech/tags/outage.md>), [outages](<https://devfeed.tech/tags/outages.md>), [recovery](<https://devfeed.tech/tags/recovery.md>), [resilience](<https://devfeed.tech/tags/resilience.md>), [systems](<https://devfeed.tech/tags/systems.md>), [testing](<https://devfeed.tech/tags/testing.md>), [uk](<https://devfeed.tech/tags/uk.md>), [verify](<https://devfeed.tech/tags/verify.md>)

### AI overview

This article explains how reliability engineering and Gremlin can test disaster recovery plans by safely simulating failures, combining faults into realistic scenarios, and running repeatable test suites. The approach helps organizations verify resilience, maintain minimum service thresholds, and demonstrate regulatory compliance.

### Source excerpt

Learn how reliability engineering and Gremlin can help test your disaster recovery plans to make sure you're prepared--and compliant with regulations.

## Five ways Gremlin helps organizations meet DORA requirements

DevFeed: [Five ways Gremlin helps organizations meet DORA requirements](<https://devfeed.tech/articles/five-ways-gremlin-helps-organizations-meet-dora-requirements-11591.md>)

Original publisher: [Read original article](<https://www.gremlin.com/blog/how-gremlin-helps-meet-dora-resilience>)

Author: Ryan Detwiller

Published: 2024-05-07T00:00:00Z

Content type: article

Language: en

Sources: [Gremlin Blog](<https://devfeed.tech/sources/gremlin-blog.md>)

Topics: [Security](<https://devfeed.tech/topics/security.md>), [Resilience](<https://devfeed.tech/topics/resilience.md>), [Disaster Recovery](<https://devfeed.tech/topics/disaster-recovery.md>), [incident](<https://devfeed.tech/topics/incident.md>), [Network](<https://devfeed.tech/topics/network.md>), [systems](<https://devfeed.tech/topics/systems.md>)

Tags: [article](<https://devfeed.tech/tags/article.md>), [disaster-recovery](<https://devfeed.tech/tags/disaster-recovery.md>), [eu](<https://devfeed.tech/tags/eu.md>), [financial-services](<https://devfeed.tech/tags/financial-services.md>), [gremlin](<https://devfeed.tech/tags/gremlin.md>), [incident](<https://devfeed.tech/tags/incident.md>), [infrastructure](<https://devfeed.tech/tags/infrastructure.md>), [monitoring](<https://devfeed.tech/tags/monitoring.md>), [network](<https://devfeed.tech/tags/network.md>), [operational](<https://devfeed.tech/tags/operational.md>), [outages](<https://devfeed.tech/tags/outages.md>), [performance](<https://devfeed.tech/tags/performance.md>), [recovery](<https://devfeed.tech/tags/recovery.md>), [reliability-management](<https://devfeed.tech/tags/reliability-management.md>), [resilience](<https://devfeed.tech/tags/resilience.md>), [systems](<https://devfeed.tech/tags/systems.md>), [technology](<https://devfeed.tech/tags/technology.md>), [testing](<https://devfeed.tech/tags/testing.md>), [use-cases](<https://devfeed.tech/tags/use-cases.md>)

### AI overview

The article explains five ways Gremlin's Reliability Management Platform helps financial organizations meet the European Union's Digital Operational Resilience Act (DORA). It focuses on automated tracking, monitoring, and testing of ICT services and infrastructure, including fault-injection scenarios, reliability tests, capacity validation, incident detection and response, disaster recovery, and business continuity planning.

### Source excerpt

DORA establishes stringent standards for financial services firms operating in the EU. Gremlin's Reliability Management Platform helps organizations meet DORA requirements by automating the tracking, monitoring, and testing of ICT services and infrastructure for resiliency risks. This article discusses five ways Gremlin can help.

## High availability and disaster recovery with Temporal Cloud

DevFeed: [High availability and disaster recovery with Temporal Cloud](<https://devfeed.tech/articles/high-availability-and-disaster-recovery-with-temporal-cloud-35847.md>)

Original publisher: [Read original article](<https://temporal.io/blog/high-availability-and-disaster-recovery-with-temporal-cloud>)

Author: Meagan Speare

Published: 2024-03-27T06:00:00Z

Content type: article

Language: en

Sources: [Temporal Blog](<https://devfeed.tech/sources/temporal-blog.md>)

Topics: [Availability](<https://devfeed.tech/topics/availability.md>), [Disaster Recovery](<https://devfeed.tech/topics/disaster-recovery.md>), [Self-hosted](<https://devfeed.tech/topics/self-hosted.md>), [distributed-systems](<https://devfeed.tech/topics/distributed-systems.md>)

Tags: [availability](<https://devfeed.tech/tags/availability.md>), [cloud](<https://devfeed.tech/tags/cloud.md>), [disaster-recovery](<https://devfeed.tech/tags/disaster-recovery.md>), [fault-tolerance](<https://devfeed.tech/tags/fault-tolerance.md>), [reliability](<https://devfeed.tech/tags/reliability.md>), [self-hosting](<https://devfeed.tech/tags/self-hosting.md>), [temporal-concepts](<https://devfeed.tech/tags/temporal-concepts.md>)

### AI overview

The article explains the operational challenges of self-hosting Temporal at scale, including maintaining multiple independently scalable services and their supporting database. It describes how Temporal Cloud provides managed high availability, fault tolerance, reliability, and disaster-recovery capabilities for mission-critical and high-scale applications.

### Source excerpt

Learn how Temporal Cloud supports high availability and disaster recovery to keep your applications reliable at scale.

[Next page](<https://devfeed.tech/topics/disaster-recovery.md?cursor=WyIyMDI0LTAzLTI3VDA2OjAwOjAwKzAwOjAwIiwgIjRjZTU5YTVkLTg2MTEtNDI3OS04NWU1LTU2YTJmNmY3NzJiOCJd>)