1-50 of 10000 results (111ms)
2026-08-19 §
02:54 <ryankemper@cumin2003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on an-worker1186.eqiad.wmnet with reason: host reimage [production]
02:46 <ryankemper@cumin2003> START - Cookbook sre.hosts.downtime for 2:00:00 on an-worker1186.eqiad.wmnet with reason: host reimage [production]
02:46 <marostegui@cumin1003> dbctl commit (dc=all): 'Depool db2207 T435270', diff saved to https://phabricator.wikimedia.org/P96181 and previous config saved to /var/cache/conftool/dbconfig/20260819-024627-marostegui.json [production]
02:44 <marostegui@cumin1003> dbctl commit (dc=all): 'Promote db2204 to s2 primary T435270', diff saved to https://phabricator.wikimedia.org/P96180 and previous config saved to /var/cache/conftool/dbconfig/20260819-024403-marostegui.json [production]
02:43 <marostegui> Starting s2 codfw failover from db2207 to db2204 - T435270 [production]
02:39 <marostegui@cumin1003> dbctl commit (dc=all): 'Set db2204 with weight 0 T435270', diff saved to https://phabricator.wikimedia.org/P96179 and previous config saved to /var/cache/conftool/dbconfig/20260819-023951-marostegui.json [production]
02:39 <marostegui@cumin1003> DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 1:00:00 on 28 hosts with reason: Primary switchover s2 T435270 [production]
02:32 <ryankemper@cumin2003> START - Cookbook sre.hosts.reimage for host an-worker1186.eqiad.wmnet with OS bookworm [production]
02:29 <ryankemper@cumin2003> END (FAIL) - Cookbook sre.hosts.reimage (exit_code=99) for host an-worker1186.eqiad.wmnet with OS bookworm [production]
02:18 <denisse@cumin1003> END (FAIL) - Cookbook sre.mysql.depool (exit_code=99) depool db2207: Depooling replica [production]
02:18 <denisse@cumin1003> START - Cookbook sre.mysql.depool depool db2207: Depooling replica [production]
02:07 <mwpresync@deploy1003> Finished scap build-images: Publishing wmf/next image (duration: 06m 48s) [production]
02:00 <mwpresync@deploy1003> Started scap build-images: Publishing wmf/next image [production]
2026-08-18 §
23:55 <krinkle@deploy1003> Finished scap sync-world: Backport for [[gerrit:1264845|Remove unused/redundant wgMFNoindexPages=true setting (T255458)]] (duration: 09m 42s) [production]
23:51 <krinkle@deploy1003> krinkle: Continuing with deployment [production]
23:47 <krinkle@deploy1003> krinkle: Backport for [[gerrit:1264845|Remove unused/redundant wgMFNoindexPages=true setting (T255458)]] synced to the testservers (see https://wikitech.wikimedia.org/wiki/Mwdebug). Changes can now be verified there. [production]
23:45 <krinkle@deploy1003> Started scap sync-world: Backport for [[gerrit:1264845|Remove unused/redundant wgMFNoindexPages=true setting (T255458)]] [production]
23:38 <ladsgroup@deploy1003> Finished scap sync-world: Backport for [[gerrit:1326944|Retire filebackend lock manager in favour of the default one (T366938)]] (duration: 08m 55s) [production]
23:34 <ladsgroup@deploy1003> ladsgroup: Continuing with deployment [production]
23:31 <ladsgroup@deploy1003> ladsgroup: Backport for [[gerrit:1326944|Retire filebackend lock manager in favour of the default one (T366938)]] synced to the testservers (see https://wikitech.wikimedia.org/wiki/Mwdebug). Changes can now be verified there. [production]
23:29 <ladsgroup@deploy1003> Started scap sync-world: Backport for [[gerrit:1326944|Retire filebackend lock manager in favour of the default one (T366938)]] [production]
23:27 <ryankemper@cumin2003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host an-worker1170.eqiad.wmnet with OS bookworm [production]
23:21 <ryankemper@cumin2003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host an-worker1205.eqiad.wmnet with OS bookworm [production]
23:20 <ryankemper@cumin2003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host an-worker1171.eqiad.wmnet with OS bookworm [production]
23:15 <ryankemper@cumin2003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host an-worker1206.eqiad.wmnet with OS bookworm [production]
23:05 <ryankemper@cumin2003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on an-worker1170.eqiad.wmnet with reason: host reimage [production]
23:00 <ryankemper@cumin2003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on an-worker1205.eqiad.wmnet with reason: host reimage [production]
22:57 <ryankemper@cumin2003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on an-worker1171.eqiad.wmnet with reason: host reimage [production]
22:54 <ryankemper@cumin2003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on an-worker1206.eqiad.wmnet with reason: host reimage [production]
22:53 <ryankemper@cumin2003> START - Cookbook sre.hosts.downtime for 2:00:00 on an-worker1205.eqiad.wmnet with reason: host reimage [production]
22:51 <ryankemper@cumin2003> START - Cookbook sre.hosts.downtime for 2:00:00 on an-worker1171.eqiad.wmnet with reason: host reimage [production]
22:51 <ryankemper@cumin2003> START - Cookbook sre.hosts.downtime for 2:00:00 on an-worker1170.eqiad.wmnet with reason: host reimage [production]
22:50 <ryankemper@cumin2003> START - Cookbook sre.hosts.downtime for 2:00:00 on an-worker1206.eqiad.wmnet with reason: host reimage [production]
22:36 <ryankemper@cumin2003> START - Cookbook sre.hosts.reimage for host an-worker1206.eqiad.wmnet with OS bookworm [production]
22:35 <ryankemper@cumin2003> START - Cookbook sre.hosts.reimage for host an-worker1205.eqiad.wmnet with OS bookworm [production]
22:35 <ryankemper@cumin2003> START - Cookbook sre.hosts.reimage for host an-worker1186.eqiad.wmnet with OS bookworm [production]
22:35 <ryankemper@cumin2003> START - Cookbook sre.hosts.reimage for host an-worker1171.eqiad.wmnet with OS bookworm [production]
22:35 <ryankemper@cumin2003> START - Cookbook sre.hosts.reimage for host an-worker1170.eqiad.wmnet with OS bookworm [production]
22:33 <ryankemper@cumin2003> END (FAIL) - Cookbook sre.hosts.reimage (exit_code=99) for host an-worker1194.eqiad.wmnet with OS bookworm [production]
22:22 <ladsgroup@deploy1003> Finished scap sync-world: Backport for [[gerrit:1326923|Enable redis lock manager on s6 (T366938)]] (duration: 11m 52s) [production]
22:18 <ladsgroup@deploy1003> ladsgroup: Continuing with deployment [production]
22:12 <ladsgroup@deploy1003> ladsgroup: Backport for [[gerrit:1326923|Enable redis lock manager on s6 (T366938)]] synced to the testservers (see https://wikitech.wikimedia.org/wiki/Mwdebug). Changes can now be verified there. [production]
22:10 <ladsgroup@deploy1003> Started scap sync-world: Backport for [[gerrit:1326923|Enable redis lock manager on s6 (T366938)]] [production]
22:04 <sbassett> Deployed security fix for T435234 (wmf.16) [production]
21:54 <sbassett> Deployed security fix for T435234 (wmf.15) [production]
21:38 <caro@deploy1003> Finished scap sync-world: Backport for [[gerrit:1326925|Exclude NPOV-type LLM-generated suggestions (T435253)]], [[gerrit:1326926|Exclude NPOV-type LLM-generated suggestions (T435253)]], [[gerrit:1326929|LLMSuggestionsEditCheck: final comparison should also have the object-replacements]], [[gerrit:1326928|LLMSuggestionsEditCheck: final comparison should also have the object-replacements]] (duration: 0 [production]
21:34 <caro@deploy1003> caro: Continuing with deployment [production]
21:33 <caro@deploy1003> caro: Backport for [[gerrit:1326925|Exclude NPOV-type LLM-generated suggestions (T435253)]], [[gerrit:1326926|Exclude NPOV-type LLM-generated suggestions (T435253)]], [[gerrit:1326929|LLMSuggestionsEditCheck: final comparison should also have the object-replacements]], [[gerrit:1326928|LLMSuggestionsEditCheck: final comparison should also have the object-replacements]] synced to the testservers (see h [production]
21:31 <caro@deploy1003> Started scap sync-world: Backport for [[gerrit:1326925|Exclude NPOV-type LLM-generated suggestions (T435253)]], [[gerrit:1326926|Exclude NPOV-type LLM-generated suggestions (T435253)]], [[gerrit:1326929|LLMSuggestionsEditCheck: final comparison should also have the object-replacements]], [[gerrit:1326928|LLMSuggestionsEditCheck: final comparison should also have the object-replacements]] [production]
21:24 <ryankemper@cumin2003> DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 4:00:00 on an-worker1204.eqiad.wmnet with reason: 1204 datanode repair T434494 [production]