1-50 of 10000 results (54ms)
2026-10-02 ยง
21:10 <jclark@cumin1004> END (PASS) - Cookbook sre.hosts.provision (exit_code=0) for host an-druid1003.mgmt.eqiad.wmnet with chassis set policy GRACEFUL_RESTART and with Dell SCP reboot policy GRACEFUL [production]
21:01 <jclark@cumin1004> START - Cookbook sre.hosts.provision for host an-druid1003.mgmt.eqiad.wmnet with chassis set policy GRACEFUL_RESTART and with Dell SCP reboot policy GRACEFUL [production]
20:54 <inflatador> bking@an-druid100[4-7] `tune2fs -m0 /dev/mapper/vg0-srv` T440033 [production]
20:53 <btullis@cumin1004> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host dse-k8s-worker2010.codfw.wmnet with OS bookworm [production]
20:49 <btullis@cumin1004> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host dse-k8s-worker2011.codfw.wmnet with OS bookworm [production]
20:40 <jclark@cumin1004> START - Cookbook sre.hosts.reimage for host an-druid1003.eqiad.wmnet with OS bookworm [production]
20:37 <btullis@cumin1004> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on dse-k8s-worker2010.codfw.wmnet with reason: host reimage [production]
20:31 <btullis@cumin1004> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on dse-k8s-worker2011.codfw.wmnet with reason: host reimage [production]
20:26 <btullis@cumin1004> START - Cookbook sre.hosts.downtime for 2:00:00 on dse-k8s-worker2010.codfw.wmnet with reason: host reimage [production]
20:26 <btullis@cumin1004> START - Cookbook sre.hosts.downtime for 2:00:00 on dse-k8s-worker2011.codfw.wmnet with reason: host reimage [production]
20:23 <osleger@deploy1003> helmfile [codfw] DONE helmfile.d/services/mw-parsoid: apply [production]
20:23 <osleger@deploy1003> helmfile [codfw] START helmfile.d/services/mw-parsoid: apply [production]
20:23 <osleger@deploy1003> helmfile [eqiad] DONE helmfile.d/services/mw-parsoid: apply [production]
20:23 <osleger@deploy1003> helmfile [eqiad] START helmfile.d/services/mw-parsoid: apply [production]
20:22 <osleger@deploy1003> helmfile [codfw] DONE helmfile.d/services/mw-parsoid: apply [production]
20:21 <osleger@deploy1003> helmfile [codfw] START helmfile.d/services/mw-parsoid: apply [production]
20:21 <osleger@deploy1003> helmfile [eqiad] DONE helmfile.d/services/mw-parsoid: apply [production]
20:20 <osleger@deploy1003> helmfile [eqiad] START helmfile.d/services/mw-parsoid: apply [production]
20:17 <dzahn@cumin1004> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host zuul1004.eqiad.wmnet with OS trixie [production]
20:10 <btullis@cumin1004> START - Cookbook sre.hosts.reimage for host dse-k8s-worker2010.codfw.wmnet with OS bookworm [production]
20:10 <btullis@cumin1004> START - Cookbook sre.hosts.reimage for host dse-k8s-worker2011.codfw.wmnet with OS bookworm [production]
19:58 <dzahn@cumin1004> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on zuul1004.eqiad.wmnet with reason: host reimage [production]
19:54 <dzahn@cumin1004> START - Cookbook sre.hosts.downtime for 2:00:00 on zuul1004.eqiad.wmnet with reason: host reimage [production]
19:47 <bking@cumin2003> START - Cookbook sre.hardware.upgrade-firmware upgrade firmware for hosts ['an-druid1003'] [production]
19:46 <bking@cumin2003> END (PASS) - Cookbook sre.hardware.upgrade-firmware (exit_code=0) upgrade firmware for hosts ['an-druid1003'] [production]
19:39 <dzahn@cumin1004> START - Cookbook sre.hosts.reimage for host zuul1004.eqiad.wmnet with OS trixie [production]
19:37 <bking@cumin2003> START - Cookbook sre.hardware.upgrade-firmware upgrade firmware for hosts ['an-druid1003'] [production]
19:37 <bking@cumin2003> END (PASS) - Cookbook sre.hardware.upgrade-firmware (exit_code=0) upgrade firmware for hosts ['an-druid1003'] [production]
19:37 <bking@cumin2003> START - Cookbook sre.hardware.upgrade-firmware upgrade firmware for hosts ['an-druid1003'] [production]
19:33 <mutante> zuul1004 - reimage stuck in d-i on leftover software RAID/LVM from its previous life as a ganeti host; wiping old md/LVM metadata and partition tables from the installer shell same fix as on zuul1005 (as also mentioned by Valerie/dcops) - T427353 [production]
19:08 <dzahn@cumin1004> END (FAIL) - Cookbook sre.hosts.reimage (exit_code=99) for host zuul1004.eqiad.wmnet with OS trixie [production]
18:40 <bking@cumin2003> END (PASS) - Cookbook sre.hosts.reboot-single (exit_code=0) for host flink-zk2003.codfw.wmnet [production]
18:36 <bking@cumin2003> START - Cookbook sre.hosts.reboot-single for host flink-zk2003.codfw.wmnet [production]
18:03 <dzahn@cumin1004> START - Cookbook sre.hosts.reimage for host zuul1004.eqiad.wmnet with OS trixie [production]
18:01 <dzahn@cumin1004> END (PASS) - Cookbook sre.hosts.decommission (exit_code=0) for hosts zuul1001.eqiad.wmnet [production]
18:01 <dzahn@cumin1004> END (PASS) - Cookbook sre.dns.netbox (exit_code=0) [production]
18:01 <dzahn@cumin1004> END (PASS) - Cookbook sre.puppet.sync-netbox-hiera (exit_code=0) generate netbox hiera data: "Triggered by cookbooks.sre.dns.netbox: zuul1001.eqiad.wmnet decommissioned, removing all IPs except the asset tag one - dzahn@cumin1004" [production]
18:01 <dzahn@cumin1004> START - Cookbook sre.puppet.sync-netbox-hiera generate netbox hiera data: "Triggered by cookbooks.sre.dns.netbox: zuul1001.eqiad.wmnet decommissioned, removing all IPs except the asset tag one - dzahn@cumin1004" [production]
18:00 <bking@cumin2003> END (PASS) - Cookbook sre.hosts.reboot-single (exit_code=0) for host flink-zk2002.codfw.wmnet [production]
18:00 <bking@cumin2003> END (PASS) - Cookbook sre.hosts.reboot-single (exit_code=0) for host flink-zk1003.eqiad.wmnet [production]
17:56 <bking@cumin2003> START - Cookbook sre.hosts.reboot-single for host flink-zk2002.codfw.wmnet [production]
17:56 <bking@cumin2003> START - Cookbook sre.hosts.reboot-single for host flink-zk1003.eqiad.wmnet [production]
17:51 <bking@cumin2003> END (PASS) - Cookbook sre.hosts.reboot-single (exit_code=0) for host flink-zk2001.codfw.wmnet [production]
17:50 <dzahn@cumin1004> START - Cookbook sre.dns.netbox [production]
17:47 <bking@cumin2003> START - Cookbook sre.hosts.reboot-single for host flink-zk2001.codfw.wmnet [production]
17:46 <bking@cumin2003> END (PASS) - Cookbook sre.hosts.reboot-single (exit_code=0) for host flink-zk1002.eqiad.wmnet [production]
17:45 <dzahn@cumin1004> START - Cookbook sre.hosts.decommission for hosts zuul1001.eqiad.wmnet [production]
17:44 <dzahn@dns1004> END - running authdns-update [production]
17:42 <bking@cumin2003> START - Cookbook sre.hosts.reboot-single for host flink-zk1002.eqiad.wmnet [production]
17:41 <dzahn@dns1004> START - running authdns-update [production]