|
2026-08-19
§
|
| 03:16 |
<ryankemper@cumin2003> |
END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host an-worker1186.eqiad.wmnet with OS bookworm |
[production] |
| 02:54 |
<ryankemper@cumin2003> |
END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on an-worker1186.eqiad.wmnet with reason: host reimage |
[production] |
| 02:51 |
<andrew@cloudcumin1001> |
END (PASS) - Cookbook wmcs.ceph.unset_cluster_maintenance (exit_code=0) |
[admin] |
| 02:51 |
<andrew@cloudcumin1001> |
START - Cookbook wmcs.ceph.unset_cluster_maintenance |
[admin] |
| 02:46 |
<ryankemper@cumin2003> |
START - Cookbook sre.hosts.downtime for 2:00:00 on an-worker1186.eqiad.wmnet with reason: host reimage |
[production] |
| 02:46 |
<marostegui@cumin1003> |
dbctl commit (dc=all): 'Depool db2207 T435270', diff saved to https://phabricator.wikimedia.org/P96181 and previous config saved to /var/cache/conftool/dbconfig/20260819-024627-marostegui.json |
[production] |
| 02:44 |
<marostegui@cumin1003> |
dbctl commit (dc=all): 'Promote db2204 to s2 primary T435270', diff saved to https://phabricator.wikimedia.org/P96180 and previous config saved to /var/cache/conftool/dbconfig/20260819-024403-marostegui.json |
[production] |
| 02:43 |
<marostegui> |
Starting s2 codfw failover from db2207 to db2204 - T435270 |
[production] |
| 02:39 |
<marostegui@cumin1003> |
dbctl commit (dc=all): 'Set db2204 with weight 0 T435270', diff saved to https://phabricator.wikimedia.org/P96179 and previous config saved to /var/cache/conftool/dbconfig/20260819-023951-marostegui.json |
[production] |
| 02:39 |
<marostegui@cumin1003> |
DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 1:00:00 on 28 hosts with reason: Primary switchover s2 T435270 |
[production] |
| 02:33 |
<eileen> |
SmashPig upgraded from 0e3e59b4 to 4e9d9df0 |
[fundraising] |
| 02:32 |
<ryankemper@cumin2003> |
START - Cookbook sre.hosts.reimage for host an-worker1186.eqiad.wmnet with OS bookworm |
[production] |
| 02:29 |
<ryankemper@cumin2003> |
END (FAIL) - Cookbook sre.hosts.reimage (exit_code=99) for host an-worker1186.eqiad.wmnet with OS bookworm |
[production] |
| 02:18 |
<denisse@cumin1003> |
END (FAIL) - Cookbook sre.mysql.depool (exit_code=99) depool db2207: Depooling replica |
[production] |
| 02:18 |
<denisse@cumin1003> |
START - Cookbook sre.mysql.depool depool db2207: Depooling replica |
[production] |
| 02:10 |
<andrew@cloudcumin1001> |
END (PASS) - Cookbook wmcs.ceph.set_cluster_in_maintenance (exit_code=0) |
[admin] |
| 02:09 |
<andrew@cloudcumin1001> |
START - Cookbook wmcs.ceph.set_cluster_in_maintenance |
[admin] |
| 02:07 |
<mwpresync@deploy1003> |
Finished scap build-images: Publishing wmf/next image (duration: 06m 48s) |
[production] |
| 02:00 |
<mwpresync@deploy1003> |
Started scap build-images: Publishing wmf/next image |
[production] |
| 00:25 |
<eileen> |
SmashPig upgraded from 89909858 to 0e3e59b4 |
[fundraising] |
|
2026-08-18
§
|
| 23:55 |
<krinkle@deploy1003> |
Finished scap sync-world: Backport for [[gerrit:1264845|Remove unused/redundant wgMFNoindexPages=true setting (T255458)]] (duration: 09m 42s) |
[production] |
| 23:51 |
<krinkle@deploy1003> |
krinkle: Continuing with deployment |
[production] |
| 23:47 |
<krinkle@deploy1003> |
krinkle: Backport for [[gerrit:1264845|Remove unused/redundant wgMFNoindexPages=true setting (T255458)]] synced to the testservers (see https://wikitech.wikimedia.org/wiki/Mwdebug). Changes can now be verified there. |
[production] |
| 23:45 |
<krinkle@deploy1003> |
Started scap sync-world: Backport for [[gerrit:1264845|Remove unused/redundant wgMFNoindexPages=true setting (T255458)]] |
[production] |
| 23:38 |
<ladsgroup@deploy1003> |
Finished scap sync-world: Backport for [[gerrit:1326944|Retire filebackend lock manager in favour of the default one (T366938)]] (duration: 08m 55s) |
[production] |
| 23:34 |
<ladsgroup@deploy1003> |
ladsgroup: Continuing with deployment |
[production] |
| 23:31 |
<ladsgroup@deploy1003> |
ladsgroup: Backport for [[gerrit:1326944|Retire filebackend lock manager in favour of the default one (T366938)]] synced to the testservers (see https://wikitech.wikimedia.org/wiki/Mwdebug). Changes can now be verified there. |
[production] |
| 23:29 |
<ladsgroup@deploy1003> |
Started scap sync-world: Backport for [[gerrit:1326944|Retire filebackend lock manager in favour of the default one (T366938)]] |
[production] |
| 23:27 |
<ryankemper@cumin2003> |
END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host an-worker1170.eqiad.wmnet with OS bookworm |
[production] |
| 23:21 |
<ryankemper@cumin2003> |
END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host an-worker1205.eqiad.wmnet with OS bookworm |
[production] |
| 23:20 |
<ryankemper@cumin2003> |
END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host an-worker1171.eqiad.wmnet with OS bookworm |
[production] |
| 23:15 |
<ryankemper@cumin2003> |
END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host an-worker1206.eqiad.wmnet with OS bookworm |
[production] |
| 23:05 |
<ryankemper@cumin2003> |
END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on an-worker1170.eqiad.wmnet with reason: host reimage |
[production] |
| 23:00 |
<ryankemper@cumin2003> |
END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on an-worker1205.eqiad.wmnet with reason: host reimage |
[production] |
| 22:57 |
<ryankemper@cumin2003> |
END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on an-worker1171.eqiad.wmnet with reason: host reimage |
[production] |
| 22:54 |
<ryankemper@cumin2003> |
END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on an-worker1206.eqiad.wmnet with reason: host reimage |
[production] |
| 22:53 |
<ryankemper@cumin2003> |
START - Cookbook sre.hosts.downtime for 2:00:00 on an-worker1205.eqiad.wmnet with reason: host reimage |
[production] |
| 22:51 |
<ryankemper@cumin2003> |
START - Cookbook sre.hosts.downtime for 2:00:00 on an-worker1171.eqiad.wmnet with reason: host reimage |
[production] |
| 22:51 |
<ryankemper@cumin2003> |
START - Cookbook sre.hosts.downtime for 2:00:00 on an-worker1170.eqiad.wmnet with reason: host reimage |
[production] |
| 22:50 |
<ryankemper@cumin2003> |
START - Cookbook sre.hosts.downtime for 2:00:00 on an-worker1206.eqiad.wmnet with reason: host reimage |
[production] |
| 22:36 |
<ryankemper@cumin2003> |
START - Cookbook sre.hosts.reimage for host an-worker1206.eqiad.wmnet with OS bookworm |
[production] |
| 22:35 |
<ryankemper@cumin2003> |
START - Cookbook sre.hosts.reimage for host an-worker1205.eqiad.wmnet with OS bookworm |
[production] |
| 22:35 |
<ryankemper@cumin2003> |
START - Cookbook sre.hosts.reimage for host an-worker1186.eqiad.wmnet with OS bookworm |
[production] |
| 22:35 |
<ryankemper@cumin2003> |
START - Cookbook sre.hosts.reimage for host an-worker1171.eqiad.wmnet with OS bookworm |
[production] |
| 22:35 |
<ryankemper@cumin2003> |
START - Cookbook sre.hosts.reimage for host an-worker1170.eqiad.wmnet with OS bookworm |
[production] |
| 22:33 |
<ryankemper@cumin2003> |
END (FAIL) - Cookbook sre.hosts.reimage (exit_code=99) for host an-worker1194.eqiad.wmnet with OS bookworm |
[production] |
| 22:22 |
<ladsgroup@deploy1003> |
Finished scap sync-world: Backport for [[gerrit:1326923|Enable redis lock manager on s6 (T366938)]] (duration: 11m 52s) |
[production] |
| 22:18 |
<ladsgroup@deploy1003> |
ladsgroup: Continuing with deployment |
[production] |
| 22:12 |
<ladsgroup@deploy1003> |
ladsgroup: Backport for [[gerrit:1326923|Enable redis lock manager on s6 (T366938)]] synced to the testservers (see https://wikitech.wikimedia.org/wiki/Mwdebug). Changes can now be verified there. |
[production] |
| 22:10 |
<ladsgroup@deploy1003> |
Started scap sync-world: Backport for [[gerrit:1326923|Enable redis lock manager on s6 (T366938)]] |
[production] |