In this post we will see a case study of a Galera Cluster migration to AWS Aurora and quick solution to the replication issue. A friend received an error in…
This post is a lab experiment learning from migration to the Percona Xtradb Cluster (Galera) and a very unexpected DEADLOCK scenario which took me back to basics. (root@localhost) [test]>insert into…
A Percona Xtradb (Galera) Cluster node may fail to join it due to many possible mistakes causing SST to fail. It could be a configuration item or purely setup requirement. In this article we will be troubleshooting step by step the SST issues faced.
2 responses to “Debugging Percona Xtradb (Galera) Cluster node startup / SST errors”
Another important one – in case the server has multiple IPs and wsrep_sst_receive_address is set to AUTO it may use the wrong one. For instance in a test configuration with Vagrant / OVM you may have two IP address, one internal on the 10.x.x.x range and other on 192.x.x.x and PXC may use the 10.x.x.x rather than the desired 192.x.x.x address.
In this case SST will fail with a message like:
150921 14:09:40 [ERROR] WSREP: gcs/src/gcs_group.cpp:long int gcs_group_handle_join_msg(gcs_group_t*, const gcs_recv_msg_t*)():717: Will never receive state. Need to abort.
Thanks for the update Nik.