151-200 of 10000 results (11ms)
2026-07-29 ยง
14:10 <cwilliams@cumin1003> DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 1 day, 0:00:00 on db1244.eqiad.wmnet with reason: Maintenance [production]
14:10 <pt1979@cumin2002> START - Cookbook sre.dns.netbox [production]
14:10 <cwilliams@cumin1003> START - Cookbook sre.mysql.update-replication [production]
14:09 <root@cumin1003> END (PASS) - Cookbook sre.mysql.pool (exit_code=0) pool db1243: Maintenance [production]
14:09 <swfrench@cumin2002> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on conf2005.codfw.wmnet with reason: host reimage [production]
14:09 <root@cumin1003> START - Cookbook sre.mysql.pool pool db1174: Maintenance [production]
14:08 <sukhe@cumin1003> END (FAIL) - Cookbook sre.loadbalancer.migrate-service-ipip (exit_code=99) for role: url_downloader@eqiad [production]
14:06 <root@cumin1003> START - Cookbook sre.mysql.pool pool db1185: Maintenance [production]
14:06 <sukhe@cumin1003> START - Cookbook sre.loadbalancer.migrate-service-ipip for role: url_downloader@eqiad [production]
14:05 <swfrench@cumin2002> START - Cookbook sre.hosts.downtime for 2:00:00 on conf2005.codfw.wmnet with reason: host reimage [production]
14:03 <cwilliams@cumin1003> dbctl commit (dc=all): 'Depooling db1174 (T431660)', diff saved to https://phabricator.wikimedia.org/P95595 and previous config saved to /var/cache/conftool/dbconfig/20260729-140309-cwilliams.json [production]
14:03 <sukhe@dns1004> END - running authdns-update [production]
14:03 <cwilliams@cumin1003> DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 1 day, 0:00:00 on db1174.eqiad.wmnet with reason: Maintenance [production]
14:02 <root@cumin1003> END (PASS) - Cookbook sre.mysql.pool (exit_code=0) pool db1170: Maintenance [production]
14:02 <root@cumin1003> START - Cookbook sre.mysql.pool pool db1218: Maintenance [production]
14:01 <sukhe@dns1004> START - running authdns-update [production]
14:00 <sukhe@dns1004> START - running authdns-update [production]
13:59 <root@cumin1003> START - Cookbook sre.mysql.pool pool db1233: Maintenance [production]
13:59 <cwilliams@cumin1003> dbctl commit (dc=all): 'Depooling db1185 (T431660)', diff saved to https://phabricator.wikimedia.org/P95592 and previous config saved to /var/cache/conftool/dbconfig/20260729-135925-cwilliams.json [production]
13:59 <cwilliams@cumin1003> DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 1 day, 0:00:00 on db1185.eqiad.wmnet with reason: Maintenance [production]
13:59 <root@cumin1003> END (PASS) - Cookbook sre.mysql.pool (exit_code=0) pool db1161: Maintenance [production]
13:58 <sukhe@puppetserver1001> conftool action : set/pooled=true; selector: dnsdisc=urldownloader [production]
13:58 <root@cumin1003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host db1265.eqiad.wmnet with OS trixie [production]
13:56 <cwilliams@cumin1003> dbctl commit (dc=all): 'Depooling db1218 (T431660)', diff saved to https://phabricator.wikimedia.org/P95590 and previous config saved to /var/cache/conftool/dbconfig/20260729-135621-cwilliams.json [production]
13:56 <cwilliams@cumin1003> DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 1 day, 0:00:00 on db1218.eqiad.wmnet with reason: Maintenance [production]
13:55 <root@cumin1003> END (PASS) - Cookbook sre.mysql.pool (exit_code=0) pool db1206: Maintenance [production]
13:54 <mvernon@cumin2003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host ms-be2073.codfw.wmnet with OS trixie [production]
13:53 <cwilliams@cumin1003> dbctl commit (dc=all): 'Depooling db1233 (T431660)', diff saved to https://phabricator.wikimedia.org/P95587 and previous config saved to /var/cache/conftool/dbconfig/20260729-135335-cwilliams.json [production]
13:53 <cwilliams@cumin1003> DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 1 day, 0:00:00 on db1233.eqiad.wmnet with reason: Maintenance [production]
13:53 <root@cumin1003> END (PASS) - Cookbook sre.mysql.pool (exit_code=0) pool db1229: Maintenance [production]
13:50 <jgiannelos@deploy1003> helmfile [eqiad] DONE helmfile.d/services/kartotherian: apply [production]
13:50 <sukhe> sukhe@lvs2013:~$ sudo systemctl restart pybal.service [production]
13:49 <blake@deploy1003> helmfile [codfw] DONE helmfile.d/services/mw-pretrain: apply [production]
13:49 <jgiannelos@deploy1003> helmfile [eqiad] START helmfile.d/services/kartotherian: apply [production]
13:48 <jgiannelos@deploy1003> helmfile [codfw] DONE helmfile.d/services/kartotherian: apply [production]
13:47 <mvernon@cumin1003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host ms-be1077.eqiad.wmnet with OS trixie [production]
13:47 <jgiannelos@deploy1003> helmfile [codfw] START helmfile.d/services/kartotherian: apply [production]
13:46 <swfrench@cumin2002> START - Cookbook sre.hosts.reimage for host conf2005.codfw.wmnet with OS bookworm [production]
13:44 <jgiannelos@deploy1003> helmfile [codfw] DONE helmfile.d/services/kartotherian: apply [production]
13:44 <root@cumin1003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on db1265.eqiad.wmnet with reason: host reimage [production]
13:40 <stran@deploy1003> Finished scap sync-world: Backport for [[gerrit:1319079|SI: Instrument interaction timeline link (T433053)]], [[gerrit:1319080|SI: Instrument interaction timeline link (T433053)]], [[gerrit:1319077|SI: Instrument abuse filter hits link (T433053)]], [[gerrit:1319078|SI: Instrument abuse filter hits link (T433053)]] (duration: 09m 22s) [production]
13:39 <blake@deploy1003> helmfile [codfw] START helmfile.d/services/mw-pretrain: apply [production]
13:38 <blake@deploy1003> helmfile [codfw] DONE helmfile.d/services/mw-pretrain: apply [production]
13:36 <root@cumin1003> START - Cookbook sre.hosts.downtime for 2:00:00 on db1265.eqiad.wmnet with reason: host reimage [production]
13:35 <stran@deploy1003> stran: Continuing with deployment [production]
13:33 <jgiannelos@deploy1003> helmfile [codfw] START helmfile.d/services/kartotherian: apply [production]
13:32 <stran@deploy1003> stran: Backport for [[gerrit:1319079|SI: Instrument interaction timeline link (T433053)]], [[gerrit:1319080|SI: Instrument interaction timeline link (T433053)]], [[gerrit:1319077|SI: Instrument abuse filter hits link (T433053)]], [[gerrit:1319078|SI: Instrument abuse filter hits link (T433053)]] synced to the testservers (see https://wikitech.wikimedia.org/wiki/Mwdebug). Changes can now be verified t [production]
13:32 <mvernon@cumin2003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on ms-be2073.codfw.wmnet with reason: host reimage [production]
13:30 <klausman@cumin1003> END (PASS) - Cookbook sre.hosts.reboot-single (exit_code=0) for host ml-build1001.eqiad.wmnet [production]
13:30 <stran@deploy1003> Started scap sync-world: Backport for [[gerrit:1319079|SI: Instrument interaction timeline link (T433053)]], [[gerrit:1319080|SI: Instrument interaction timeline link (T433053)]], [[gerrit:1319077|SI: Instrument abuse filter hits link (T433053)]], [[gerrit:1319078|SI: Instrument abuse filter hits link (T433053)]] [production]