An insurance services provider discovered during a DR test that their OpenSearch snapshots would not restore cleanly. Several indices came back red and the restore aborted, which meant their documented recovery plan did not work.
The snapshot repository had been reconfigured at some point without a repository verification step, so a portion of the segment files were unreachable from the current bucket path. Restores appeared to start normally and then failed on specific shards. Compounding it, the target domain ran a newer engine version than several of the older snapshots, so some indices were outside the supported restore range and could not be restored directly at all.
Managed OpenSearch domain on AWS with snapshots stored in object storage, multi-year retention on audit indices.
We treated the repository as suspect first and verified it before attempting any more restores. Once we knew which snapshots were intact, we reconstructed the recoverable set, restored the too-old indices through an intermediate-version path, and then rebuilt the repository configuration so verification runs as part of the snapshot cycle rather than only during a DR test.
All audit-critical indices were recovered and the domain returned to green. The customer now has a restore path that has been exercised end to end, and snapshot verification failures page the on-call rather than sitting silent until the next DR test.
24/7 enterprise support for OpenSearch clusters carrying regulated search and audit workloads, including security plugin and upgrade coverage.
Assessment of the technical and licensing implications of moving a large Elasticsearch estate to OpenSearch, including client and plugin compatibility.
Whether you need architecture advisory, 24/7 support, or full managed services, AceMQ has the expertise to help.