451-500 of 10000 results (146ms)
2026-07-29 ยง
13:13 <klausman@cumin1003> END (PASS) - Cookbook sre.k8s.pool-depool-node (exit_code=0) pool for host ml-serve1014.eqiad.wmnet [production]
13:13 <klausman@cumin1003> START - Cookbook sre.k8s.pool-depool-node pool for host ml-serve1014.eqiad.wmnet [production]
13:12 <root@cumin1003> END (PASS) - Cookbook sre.mysql.pool (exit_code=0) pool db1223: Maintenance [production]
13:12 <root@cumin1003> START - Cookbook sre.mysql.pool pool db1161: Maintenance [production]
13:10 <samtar@deploy1003> anzx, samtar: Continuing with deployment [production]
13:09 <samtar@deploy1003> anzx, samtar: Backport for [[gerrit:1319075|remove throttle exceptions for concluded events]] synced to the testservers (see https://wikitech.wikimedia.org/wiki/Mwdebug). Changes can now be verified there. [production]
13:09 <root@cumin1003> START - Cookbook sre.mysql.pool pool db1206: Maintenance [production]
13:08 <sukhe> sukhe@lvs2014:~$ sudo systemctl restart pybal.service [production]
13:07 <samtar@deploy1003> Started scap sync-world: Backport for [[gerrit:1319075|remove throttle exceptions for concluded events]] [production]
13:07 <cwilliams@cumin1003> dbctl commit (dc=all): 'Depooling db1170 (T431660)', diff saved to https://phabricator.wikimedia.org/P95566 and previous config saved to /var/cache/conftool/dbconfig/20260729-130730-cwilliams.json [production]
13:07 <cwilliams@cumin1003> DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 1 day, 0:00:00 on db1170.eqiad.wmnet with reason: Maintenance [production]
13:07 <cmooney@cumin1003> END (PASS) - Cookbook sre.dns.netbox (exit_code=0) [production]
13:07 <cmooney@cumin1003> END (PASS) - Cookbook sre.puppet.sync-netbox-hiera (exit_code=0) generate netbox hiera data: "Triggered by cookbooks.sre.dns.netbox: add mgmt IPs new switches - cmooney@cumin1003" [production]
13:07 <root@cumin1003> END (PASS) - Cookbook sre.mysql.pool (exit_code=0) pool db1158: Maintenance [production]
13:06 <klausman@cumin1003> END (PASS) - Cookbook sre.k8s.pool-depool-node (exit_code=0) depool for host ml-serve1014.eqiad.wmnet [production]
13:06 <cwilliams@cumin1003> dbctl commit (dc=all): 'Depooling db1161 (T431660)', diff saved to https://phabricator.wikimedia.org/P95564 and previous config saved to /var/cache/conftool/dbconfig/20260729-130616-cwilliams.json [production]
13:06 <cwilliams@cumin1003> DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2 days, 0:00:00 on 6 hosts with reason: Maintenance [production]
13:06 <mvernon@cumin1003> START - Cookbook sre.hosts.reimage for host ms-be1077.eqiad.wmnet with OS trixie [production]
13:05 <root@cumin1003> START - Cookbook sre.mysql.pool pool db1229: Maintenance [production]
13:05 <cwilliams@cumin1003> DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 1 day, 0:00:00 on db1161.eqiad.wmnet with reason: Maintenance [production]
13:05 <cmooney@cumin1003> START - Cookbook sre.puppet.sync-netbox-hiera generate netbox hiera data: "Triggered by cookbooks.sre.dns.netbox: add mgmt IPs new switches - cmooney@cumin1003" [production]
13:05 <mvernon@cumin2003> START - Cookbook sre.hosts.reimage for host ms-be2073.codfw.wmnet with OS trixie [production]
13:05 <root@cumin1003> END (PASS) - Cookbook sre.mysql.pool (exit_code=0) pool db1159: Maintenance [production]
13:02 <cwilliams@cumin1003> dbctl commit (dc=all): 'Depooling db1206 (T431660)', diff saved to https://phabricator.wikimedia.org/P95562 and previous config saved to /var/cache/conftool/dbconfig/20260729-130258-cwilliams.json [production]
13:02 <cwilliams@cumin1003> DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 1 day, 0:00:00 on db1206.eqiad.wmnet with reason: Maintenance [production]
13:02 <root@cumin1003> END (PASS) - Cookbook sre.mysql.pool (exit_code=0) pool db1196: Maintenance [production]
13:01 <sukhe> sukhe@lvs1020:~$ sudo systemctl restart pybal.service [production]
13:01 <cmooney@cumin1003> START - Cookbook sre.dns.netbox [production]
13:00 <jgiannelos@deploy1003> helmfile [staging] DONE helmfile.d/services/kartotherian: apply [production]
13:00 <jgiannelos@deploy1003> helmfile [staging] START helmfile.d/services/kartotherian: apply [production]
12:59 <cwilliams@cumin1003> dbctl commit (dc=all): 'Depooling db1229 (T431660)', diff saved to https://phabricator.wikimedia.org/P95559 and previous config saved to /var/cache/conftool/dbconfig/20260729-125950-cwilliams.json [production]
12:59 <cwilliams@cumin1003> DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 1 day, 0:00:00 on db1229.eqiad.wmnet with reason: Maintenance [production]
12:59 <root@cumin1003> END (PASS) - Cookbook sre.mysql.pool (exit_code=0) pool db1222: Maintenance [production]
12:57 <sukhe> sudo cumin 'A:lvs and (A:eqiad or A:codfw)' 'disable-puppet "adding new service urldownloader"': T429175 [production]
12:56 <klausman@cumin1003> START - Cookbook sre.k8s.pool-depool-node depool for host ml-serve1014.eqiad.wmnet [production]
12:56 <klausman@cumin1003> END (PASS) - Cookbook sre.k8s.pool-depool-node (exit_code=0) pool for host ml-serve1013.eqiad.wmnet [production]
12:56 <klausman@cumin1003> START - Cookbook sre.k8s.pool-depool-node pool for host ml-serve1013.eqiad.wmnet [production]
12:50 <klausman@cumin1003> END (PASS) - Cookbook sre.k8s.pool-depool-node (exit_code=0) depool for host ml-serve1013.eqiad.wmnet [production]
12:50 <sukhe> sudo cumin 'O:url_downloader' 'run-puppet-agent --enable "merging CR 1313948"': T429175 [production]
12:48 <btullis@cumin1003> START - Cookbook sre.puppet.sync-netbox-hiera generate netbox hiera data: "Triggered by cookbooks.sre.dns.netbox: an-test-master[1001-1002].eqiad.wmnet decommissioned, removing all IPs except the asset tag one - btullis@cumin1003" [production]
12:45 <klausman@cumin1003> START - Cookbook sre.k8s.pool-depool-node depool for host ml-serve1013.eqiad.wmnet [production]
12:45 <klausman@cumin1003> END (PASS) - Cookbook sre.k8s.pool-depool-node (exit_code=0) pool for host ml-serve1012.eqiad.wmnet [production]
12:45 <klausman@cumin1003> START - Cookbook sre.k8s.pool-depool-node pool for host ml-serve1012.eqiad.wmnet [production]
12:45 <sukhe> sudo cumin 'O:url_downloader' 'disable-puppet "merging CR 1313948"': T429175 [production]
12:44 <btullis@cumin1003> START - Cookbook sre.dns.netbox [production]
12:40 <cwilliams@cumin1003> END (FAIL) - Cookbook sre.mysql.multiinstance_reboot (exit_code=99) for db-test2001.codfw.wmnet [production]
12:40 <cwilliams@cumin1003> START - Cookbook sre.mysql.multiinstance_reboot for db-test2001.codfw.wmnet [production]
12:38 <ayounsi@dns1004> END - running authdns-update [production]
12:37 <klausman@cumin1003> END (PASS) - Cookbook sre.k8s.pool-depool-node (exit_code=0) depool for host ml-serve1012.eqiad.wmnet [production]
12:35 <ayounsi@dns1004> START - running authdns-update [production]