# Only able to start first node in cluster

**URL:** <https://forums.percona.com/t/only-able-to-start-first-node-in-cluster/3139>\
**Category:** Percona XtraDB Cluster 5.x\
**Created:** [December 6, 2013, 5:14am UTC](https://forums.percona.com/t/only-able-to-start-first-node-in-cluster/3139 "2013-12-06T05:14:37Z")\
**Posts on this page:** 2\
**Page:** 1

<div class="post-metadata">

**Author:** ![kharris](https://avatars.discourse-cdn.com/v4/letter/k/b77776/32.png) [@kharris](https://forums.percona.com/u/kharris)\
**Post date:** [December 6, 2013, 5:14am UTC](https://forums.percona.com/t/only-able-to-start-first-node-in-cluster/3139/1 "2013-12-06T05:14:37Z")

</div>

Hello,

I’m having a bit of trouble with my Percona XtraDB cluster. I’m not quite sure what started this as it was working just fine. I had to reboot one of the servers (there are 3 nodes) and it wouldn’t start again. Then I restarted a second and it too wouldn’t restart. Currently the only one that runs is the one I bootscrap by using the command /etc/init.d/mysql bootstrap-pxc. Below are entries from the log file and other details that I hope will help shed some light on this issue.

Some details that might help:

- rsync is the same version on all servers, v3.0.6
- Percona versin install from rpm: Percona-XtraDB-Cluster-shared-55-5.5.34-25.9.607

Can anyone assite me in troubleshooting this?

P.S.  
Yep, I will certainly change the password after I get this working again 🙂

Details:

mysql.cnf contents on all servers:  
[mysqld]  
wsrep\_provider=/usr/lib/libgalera\_smm.so  
wsrep\_cluster\_address=gcomm://192.168.10.31,192.168.10.32,192.168.10.39  
wsrep\_slave\_threads=8  
wsrep\_sst\_method=rsync  
binlog\_format=ROW  
default\_storage\_engine=InnoDB  
innodb\_locks\_unsafe\_for\_binlog=1  
innodb\_autoinc\_lock\_mode=2  
wsrep\_sst\_auth=sstuser:s3cret  
wsrep\_replicate\_myisam=on  
log-error=/var/log/mysqld.log

From the client log that won’t start:  
131206 04:20:54 mysqld\_safe Starting mysqld daemon with databases from /var/lib/mysql  
131206 04:20:54 mysqld\_safe WSREP: Running position recovery with --log\_error=‘/var/lib/mysql/wsrep\_recovery.7Y8asL’ --pid-file=‘/var/lib/mysql/mail3.devdigital.com-recover.pid’  
131206 04:20:57 mysqld\_safe WSREP: Recovered position 00000000-0000-0000-0000-000000000000:-1  
131206 4:20:57 [Note] WSREP: wsrep\_start\_position var submitted: ‘00000000-0000-0000-0000-000000000000:-1’  
131206 4:20:57 [Note] WSREP: Read nil XID from storage engines, skipping position init  
131206 4:20:57 [Note] WSREP: wsrep\_load(): loading provider library ‘/usr/lib/libgalera\_smm.so’  
131206 4:20:57 [Note] WSREP: wsrep\_load(): Galera 2.8(r165) by Codership Oy \<info@codership.com\> loaded successfully.  
131206 4:20:57 [Note] WSREP: Found saved state: 00000000-0000-0000-0000-000000000000:-1  
131206 4:20:57 [Note] WSREP: Reusing existing ‘/var/lib/mysql//galera.cache’.  
131206 4:20:57 [Note] WSREP: Passing config to GCS: base\_host = 192.168.10.39; base\_port = 4567; cert.log\_conflicts = no; gcache.dir = /var/lib/mysql/; gcache.keep\_pages\_size = 0; gcache.mem\_size = 0; gcache.name = /var/lib/mysql//galera.cache; gcache.page\_size = 128M; gcache.size = 128M; gcs.fc\_debug = 0; gcs.fc\_factor = 1; gcs.fc\_limit = 16; gcs.fc\_master\_slave = NO; gcs.max\_packet\_size = 64500; gcs.max\_throttle = 0.25; gcs.recv\_q\_hard\_limit = 2147483647; gcs.recv\_q\_soft\_limit = 0.25; gcs.sync\_donor = NO; replicator.causal\_read\_timeout = PT30S; replicator.commit\_order = 3  
131206 4:20:57 [Note] WSREP: Assign initial position for certification: -1, protocol version: -1  
131206 4:20:57 [Note] WSREP: wsrep\_sst\_grab()  
131206 4:20:57 [Note] WSREP: Start replication  
131206 4:20:57 [Note] WSREP: Setting initial position to 00000000-0000-0000-0000-000000000000:-1  
131206 4:20:57 [Note] WSREP: protonet asio version 0  
131206 4:20:57 [Note] WSREP: backend: asio  
131206 4:20:57 [Note] WSREP: GMCast version 0  
131206 4:20:57 [Note] WSREP: (1880d3ce-5e60-11e3-bf16-efb486db5fe4, ‘tcp://0.0.0.0:4567’) listening at tcp://0.0.0.0:4567  
131206 4:20:57 [Note] WSREP: (1880d3ce-5e60-11e3-bf16-efb486db5fe4, ‘tcp://0.0.0.0:4567’) multicast: , ttl: 1  
131206 4:20:57 [Note] WSREP: EVS version 0  
131206 4:20:57 [Note] WSREP: PC version 0  
131206 4:20:57 [Note] WSREP: gcomm: connecting to group ‘my\_wsrep\_cluster’, peer ‘192.168.10.31:,192.168.10.32:,192.168.10.39:’  
131206 4:20:57 [Warning] WSREP: (1880d3ce-5e60-11e3-bf16-efb486db5fe4, ‘tcp://0.0.0.0:4567’) address ‘tcp://192.168.10.39:4567’ points to own listening address, blacklisting  
131206 4:20:57 [Note] WSREP: (1880d3ce-5e60-11e3-bf16-efb486db5fe4, ‘tcp://0.0.0.0:4567’) address ‘tcp://192.168.10.39:4567’ pointing to uuid 1880d3ce-5e60-11e3-bf16-efb486db5fe4 is blacklisted, skipping  
131206 4:20:57 [Note] WSREP: declaring 20d04133-5e60-11e3-8215-6772cc3565ab stable  
131206 4:20:57 [Note] WSREP: Node 20d04133-5e60-11e3-8215-6772cc3565ab state prim  
131206 4:20:57 [Note] WSREP: view(view\_id(PRIM,1880d3ce-5e60-11e3-bf16-efb486db5fe4,2) memb {  
1880d3ce-5e60-11e3-bf16-efb486db5fe4,  
20d04133-5e60-11e3-8215-6772cc3565ab,  
} joined {  
} left {  
} partitioned {  
})  
131206 4:20:57 [Note] WSREP: discarding pending addr without UUID: tcp://192.168.10.32:4567  
131206 4:20:58 [Note] WSREP: gcomm: connected  
131206 4:20:58 [Note] WSREP: Changing maximum packet size to 64500, resulting msg size: 32636  
131206 4:20:58 [Note] WSREP: Shifting CLOSED → OPEN (TO: 0)  
131206 4:20:58 [Note] WSREP: Opened channel ‘my\_wsrep\_cluster’  
131206 4:20:58 [Note] WSREP: Waiting for SST to complete.  
131206 4:20:58 [Note] WSREP: New COMPONENT: primary = yes, bootstrap = no, my\_idx = 0, memb\_num = 2  
131206 4:20:58 [Note] WSREP: STATE\_EXCHANGE: sent state UUID: 191adee8-5e60-11e3-bc41-7ee12e5406fe  
131206 4:20:58 [Note] WSREP: STATE EXCHANGE: sent state msg: 191adee8-5e60-11e3-bc41-7ee12e5406fe  
131206 4:20:58 [Note] WSREP: STATE EXCHANGE: got state msg: 191adee8-5e60-11e3-bc41-7ee12e5406fe from 0 ([mail3.devdigital.com](http://mail3.devdigital.com))  
131206 4:20:58 [Note] WSREP: STATE EXCHANGE: got state msg: 191adee8-5e60-11e3-bc41-7ee12e5406fe from 1 ([mail1.devdigital.com](http://mail1.devdigital.com))  
131206 4:20:58 [Note] WSREP: Quorum results:  
version = 2,  
component = PRIMARY,  
conf\_id = 1,  
members = 1/2 (joined/total),  
act\_id = 47556,  
last\_appl. = -1,  
protocols = 0/4/2 (gcs/repl/appl),  
group UUID = 6ebd7a54-51dd-11e3-bb4c-12d4c65d64a6  
131206 4:20:58 [Note] WSREP: Flow-control interval: [23, 23]  
131206 4:20:58 [Note] WSREP: Shifting OPEN → PRIMARY (TO: 47556)  
131206 4:20:58 [Note] WSREP: State transfer required:  
Group state: 6ebd7a54-51dd-11e3-bb4c-12d4c65d64a6:47556  
Local state: 00000000-0000-0000-0000-000000000000:-1  
131206 4:20:58 [Note] WSREP: New cluster view: global state: 6ebd7a54-51dd-11e3-bb4c-12d4c65d64a6:47556, view# 2: Primary, number of nodes: 2, my index: 0, protocol version 2  
131206 4:20:58 [Warning] WSREP: Gap in state sequence. Need state transfer.  
131206 4:21:00 [Note] WSREP: Running: ‘wsrep\_sst\_rsync --role ‘joiner’ --address ‘192.168.10.39’ --auth ‘sstuser:s3cret’ --datadir ‘/var/lib/mysql/’ --defaults-file ‘/etc/my.cnf’ --parent ‘11723’’  
131206 4:21:00 [Note] WSREP: Prepared SST request: rsync|192.168.10.39:4444/rsync\_sst  
131206 4:21:00 [Note] WSREP: wsrep\_notify\_cmd is not defined, skipping notification.  
131206 4:21:00 [Note] WSREP: Assign initial position for certification: 47556, protocol version: 2  
131206 4:21:00 [Warning] WSREP: Failed to prepare for incremental state transfer: Local state UUID (00000000-0000-0000-0000-000000000000) does not match group state UUID (6ebd7a54-51dd-11e3-bb4c-12d4c65d64a6): 1 (Operation not permitted)  
at galera/src/replicator\_str.cpp:prepare\_for\_IST():445. IST will be unavailable.  
131206 4:21:00 [Note] WSREP: Node 0 ([mail3.devdigital.com](http://mail3.devdigital.com)) requested state transfer from ‘_any_’. Selected 1 ([mail1.devdigital.com](http://mail1.devdigital.com))(SYNCED) as donor.  
131206 4:21:00 [Note] WSREP: Shifting PRIMARY → JOINER (TO: 47556)  
131206 4:21:00 [Note] WSREP: Requesting state transfer: success, donor: 1  
131206 4:22:03 [Warning] WSREP: 1 ([mail1.devdigital.com](http://mail1.devdigital.com)): State transfer to 0 ([mail3.devdigital.com](http://mail3.devdigital.com)) failed: -1 (Operation not permitted)  
131206 4:22:03 [ERROR] WSREP: gcs/src/gcs\_group.c:gcs\_group\_handle\_join\_msg():719: Will never receive state. Need to abort.  
131206 4:22:03 [Note] WSREP: gcomm: terminating thread  
131206 4:22:03 [Note] WSREP: gcomm: joining thread  
131206 4:22:03 [Note] WSREP: gcomm: closing backend  
131206 4:22:04 [Note] WSREP: view(view\_id(NON\_PRIM,1880d3ce-5e60-11e3-bf16-efb486db5fe4,2) memb {  
1880d3ce-5e60-11e3-bf16-efb486db5fe4,  
} joined {  
} left {  
} partitioned {  
20d04133-5e60-11e3-8215-6772cc3565ab,  
})  
131206 4:22:04 [Note] WSREP: view((empty))  
131206 4:22:04 [Note] WSREP: gcomm: closed  
131206 4:22:04 [Note] WSREP: /usr/sbin/mysqld: Terminated.  
131206 04:22:04 mysqld\_safe mysqld from pid file /var/lib/mysql/mail3.devdigital.com.pid ended  
WSREP\_SST: [ERROR] Parent mysqld process (PID:11723) terminated unexpectedly. (20131206 04:22:04.988)  
WSREP\_SST: [INFO] Joiner cleanup. (20131206 04:22:04.992)  
WSREP\_SST: [INFO] Joiner cleanup done. (20131206 04:22:05.506)

---

<div class="post-metadata">

**Author:** ![madhusudan](https://avatars.discourse-cdn.com/v4/letter/m/76d3ee/32.png) [@madhusudan](https://forums.percona.com/u/madhusudan)\
**Post date:** [December 10, 2013, 1:47am UTC](https://forums.percona.com/t/only-able-to-start-first-node-in-cluster/3139/2 "2013-12-10T01:47:46Z")

</div>

While bootstrapping keep gcomm address empty(later you can change it). From the above log the state UUID is [COLOR=#252C2F] 00000000-0000-0000-0000-000000000000, that means its trying SST not IST.  
Try to up other nodes, and check what logs is coming.

[COLOR=#252C2F]
