Files
Tim Visée e8c8eead24 Fix getting stuck on reshard down abort (#10205)
* test: parameterize collection fixture and share resharding consensus stub

Let integration tests build a collection with a custom optimizers
config, and move NoopReshardingConsensus from the consensus idempotency
test into the shared test module.

* test: reproduce scale-down resharding abort hanging on deferred points

abort_resharding holds the shard holder write lock while
scale_down_cleanup_points deletes migrated points with
WaitUntil::Visible and no timeout. With prevent_unoptimized enabled that
wait only resolves once the optimizer has cleared every deferred point
of the shard, so a stalled optimizer wedges the write lock, and with it
every shard holder reader and - in a cluster - the consensus apply
thread driving the abort (SetShardReplicaState(Dead) on a
ReshardingScaleDown replica, e.g. after that peer is killed).

The test drives such an abort with the optimizer disabled and asserts it
completes. It currently fails by hanging into its 30s timeout, and must
pass once the cleanup delete no longer waits for visibility.

* When cleaning up old points, don't wait until visible
2026-08-12 17:10:23 +02:00

16 lines
381 B
Rust

mod collection_restore_test;
mod collection_test;
mod common;
mod consensus_idempotency_test;
mod continuous_snapshot_test;
mod create_shard_key_test;
mod distance_matrix_test;
mod grouping_test;
mod lookup_test;
mod multi_vec_test;
mod pagination_test;
mod resharding_abort_deferred_points_test;
mod snapshot_recovery_test;
mod sparse_idf_test;
mod wal_less_snapshot_clocks_test;