1-50 of 10000 results (55ms)
2026-09-10 ยง
12:05 <marostegui@cumin1003> DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 12:00:00 on clouddb1025.eqiad.wmnet with reason: Cloning x4 [production]
12:01 <btullis@cumin1004> END (PASS) - Cookbook sre.hosts.reboot-single (exit_code=0) for host dse-k8s-worker2005.codfw.wmnet [production]
11:55 <btullis@cumin1004> START - Cookbook sre.hosts.reboot-single for host dse-k8s-worker2005.codfw.wmnet [production]
11:54 <btullis@cumin1004> DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 4:00:00 on an-worker1144.eqiad.wmnet with reason: Upgrading RAID firmware [production]
11:52 <cgoubert@dns1004> END - running authdns-update [production]
11:49 <cgoubert@dns1004> START - running authdns-update [production]
11:31 <btullis@cumin1004> END (PASS) - Cookbook sre.hosts.reboot-single (exit_code=0) for host dse-k8s-worker2004.codfw.wmnet [production]
11:25 <btullis@cumin1004> START - Cookbook sre.hosts.reboot-single for host dse-k8s-worker2004.codfw.wmnet [production]
11:24 <btullis@cumin1004> DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 4:00:00 on an-worker1204.eqiad.wmnet with reason: Upgrading RAID firmware [production]
11:16 <btullis@cumin1004> END (PASS) - Cookbook sre.hosts.remove-downtime (exit_code=0) for an-worker1200.eqiad.wmnet [production]
11:16 <btullis@cumin1004> START - Cookbook sre.hosts.remove-downtime for an-worker1200.eqiad.wmnet [production]
11:04 <btullis@cumin1004> DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 4:00:00 on an-worker1200.eqiad.wmnet with reason: Upgrading RAID firmware [production]
11:03 <btullis@cumin1004> END (PASS) - Cookbook sre.hosts.remove-downtime (exit_code=0) for an-worker1199.eqiad.wmnet [production]
11:03 <btullis@cumin1004> START - Cookbook sre.hosts.remove-downtime for an-worker1199.eqiad.wmnet [production]
10:50 <sfaci@deploy1003> helmfile [dse-k8s-eqiad] DONE helmfile.d/dse-k8s-services/test-kitchen: apply [production]
10:50 <sfaci@deploy1003> helmfile [dse-k8s-eqiad] START helmfile.d/dse-k8s-services/test-kitchen: apply [production]
10:42 <btullis@cumin1004> DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 4:00:00 on an-worker1199.eqiad.wmnet with reason: Upgrading RAID firmware [production]
10:10 <marostegui@cumin1003> DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 12:00:00 on clouddb1024.eqiad.wmnet with reason: Cloning x4 [production]
10:00 <filippo@cumin1003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host cloudvirt1073.eqiad.wmnet with OS trixie [production]
09:56 <filippo@cumin1003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host cloudvirt1075.eqiad.wmnet with OS trixie [production]
09:53 <marostegui@cumin1003> conftool action : set/pooled=no; selector: name=clouddb1024.eqiad.wmnet [production]
09:45 <filippo@cumin1003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host cloudvirt1072.eqiad.wmnet with OS trixie [production]
09:41 <filippo@cumin1003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host cloudvirt1066.eqiad.wmnet with OS trixie [production]
09:37 <filippo@cumin1003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host cloudvirt1074.eqiad.wmnet with OS trixie [production]
09:24 <sfaci@deploy1003> helmfile [dse-k8s-eqiad] DONE helmfile.d/dse-k8s-services/test-kitchen: apply [production]
09:24 <sfaci@deploy1003> helmfile [dse-k8s-eqiad] START helmfile.d/dse-k8s-services/test-kitchen: apply [production]
09:04 <filippo@cumin1003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host cloudvirt1067.eqiad.wmnet with OS trixie [production]
08:51 <filippo@cumin1003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on cloudvirt1075.eqiad.wmnet with reason: host reimage [production]
08:46 <filippo@cumin1003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on cloudvirt1067.eqiad.wmnet with reason: host reimage [production]
08:45 <sfaci@deploy1003> helmfile [dse-k8s-eqiad] DONE helmfile.d/dse-k8s-services/test-kitchen-next: apply [production]
08:44 <sfaci@deploy1003> helmfile [dse-k8s-eqiad] START helmfile.d/dse-k8s-services/test-kitchen-next: apply [production]
08:44 <marostegui@cumin1003> DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 12:00:00 on db1155.eqiad.wmnet with reason: Cloning x4 [production]
08:43 <filippo@cumin1003> START - Cookbook sre.hosts.downtime for 2:00:00 on cloudvirt1075.eqiad.wmnet with reason: host reimage [production]
08:41 <filippo@cumin1003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on cloudvirt1073.eqiad.wmnet with reason: host reimage [production]
08:37 <filippo@cumin1003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on cloudvirt1074.eqiad.wmnet with reason: host reimage [production]
08:34 <mszwarc@deploy1003> Finished scap sync-world: Backport for [[gerrit:1338705|UIC: Fix getOpenContext when UserCardButton.vue is clicked (T435299)]] (duration: 09m 56s) [production]
08:33 <filippo@cumin1003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on cloudvirt1072.eqiad.wmnet with reason: host reimage [production]
08:30 <mszwarc@deploy1003> mszwarc: Continuing with deployment [production]
08:29 <filippo@cumin1003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on cloudvirt1066.eqiad.wmnet with reason: host reimage [production]
08:29 <mszwarc@deploy1003> mszwarc: Backport for [[gerrit:1338705|UIC: Fix getOpenContext when UserCardButton.vue is clicked (T435299)]] synced to the testservers (see https://wikitech.wikimedia.org/wiki/Mwdebug). Changes can now be verified there. [production]
08:28 <filippo@cumin1003> START - Cookbook sre.hosts.downtime for 2:00:00 on cloudvirt1074.eqiad.wmnet with reason: host reimage [production]
08:27 <filippo@cumin1003> START - Cookbook sre.hosts.reimage for host cloudvirt1075.eqiad.wmnet with OS trixie [production]
08:27 <filippo@cumin1003> START - Cookbook sre.hosts.downtime for 2:00:00 on cloudvirt1073.eqiad.wmnet with reason: host reimage [production]
08:25 <filippo@cumin1003> START - Cookbook sre.hosts.downtime for 2:00:00 on cloudvirt1072.eqiad.wmnet with reason: host reimage [production]
08:24 <mszwarc@deploy1003> Started scap sync-world: Backport for [[gerrit:1338705|UIC: Fix getOpenContext when UserCardButton.vue is clicked (T435299)]] [production]
08:23 <brouberol@deploy1003> helmfile [dse-k8s-eqiad] DONE helmfile.d/dse-k8s-services/pageview-trending-relative-next: apply [production]
08:23 <brouberol@deploy1003> helmfile [dse-k8s-eqiad] START helmfile.d/dse-k8s-services/pageview-trending-relative-next: apply [production]
08:23 <filippo@cumin1003> START - Cookbook sre.hosts.downtime for 2:00:00 on cloudvirt1067.eqiad.wmnet with reason: host reimage [production]
08:23 <filippo@cumin1003> START - Cookbook sre.hosts.downtime for 2:00:00 on cloudvirt1066.eqiad.wmnet with reason: host reimage [production]
08:11 <filippo@cumin1003> START - Cookbook sre.hosts.reimage for host cloudvirt1074.eqiad.wmnet with OS trixie [production]