151-200 of 10000 results (19ms)
2026-08-20 ยง
14:13 <klausman@cumin1003> END (PASS) - Cookbook sre.k8s.reboot-nodes (exit_code=0) rolling reboot on A:ml-staging-worker [production]
14:13 <klausman@cumin1003> END (PASS) - Cookbook sre.k8s.pool-depool-node (exit_code=0) pool for host ml-staging2003.codfw.wmnet [production]
14:13 <klausman@cumin1003> START - Cookbook sre.k8s.pool-depool-node pool for host ml-staging2003.codfw.wmnet [production]
14:08 <klausman@cumin1003> END (PASS) - Cookbook sre.k8s.pool-depool-node (exit_code=0) depool for host ml-staging2003.codfw.wmnet [production]
14:06 <moritzm> installing libheif security updates [production]
13:58 <klausman@cumin1003> START - Cookbook sre.k8s.pool-depool-node depool for host ml-staging2003.codfw.wmnet [production]
13:58 <klausman@cumin1003> END (PASS) - Cookbook sre.k8s.pool-depool-node (exit_code=0) pool for host ml-staging2002.codfw.wmnet [production]
13:58 <klausman@cumin1003> START - Cookbook sre.k8s.pool-depool-node pool for host ml-staging2002.codfw.wmnet [production]
13:58 <andrew@cloudcumin1001> START - Cookbook wmcs.ceph.osd.depool_and_destroy (T429387) [admin]
13:56 <cmooney@cumin1003> START - Cookbook sre.dns.netbox [production]
13:56 <fnegri@cumin1003> END (PASS) - Cookbook sre.hosts.remove-downtime (exit_code=0) for clouddb1025.eqiad.wmnet [production]
13:56 <fnegri@cumin1003> START - Cookbook sre.hosts.remove-downtime for clouddb1025.eqiad.wmnet [production]
13:56 <Lucas_WMDE> UTC afternoon backport+config window done [production]
13:53 <moritzm> installing apr-util security updates [production]
13:51 <klausman@cumin1003> END (PASS) - Cookbook sre.k8s.pool-depool-node (exit_code=0) depool for host ml-staging2002.codfw.wmnet [production]
13:50 <fnegri@cumin1003> conftool action : set/weight=100; selector: name=clouddb1025.eqiad.wmnet [production]
13:49 <fnegri@cumin1003> conftool action : set/pooled=yes; selector: name=clouddb1025.eqiad.wmnet [production]
13:42 <wmbot~jeanfred@tools-bastion-15> Deploy 881928b (Increase webservice memory limit to 1Gi) [tools.integraality]
13:42 <wmbot~jeanfred@tools-bastion-15> Deploy ea800c5 (Add functional SPARQL tests against live endpoints) [tools.integraality]
13:41 <klausman@cumin1003> START - Cookbook sre.k8s.pool-depool-node depool for host ml-staging2002.codfw.wmnet [production]
13:41 <klausman@cumin1003> END (PASS) - Cookbook sre.k8s.pool-depool-node (exit_code=0) pool for host ml-staging2001.codfw.wmnet [production]
13:41 <klausman@cumin1003> START - Cookbook sre.k8s.pool-depool-node pool for host ml-staging2001.codfw.wmnet [production]
13:41 <dpogorzelski@deploy1003> helmfile [ml-staging-codfw] DONE helmfile.d/admin 'sync'. [production]
13:38 <dpogorzelski@deploy1003> helmfile [ml-staging-codfw] START helmfile.d/admin 'sync'. [production]
13:38 <fnegri@cumin1003> DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 1:00:00 on clouddb1025.eqiad.wmnet with reason: Removing s6 from clouddb1025 [production]
13:35 <James_F> Zuul: [mediawiki/extensions/WikiLambda] Re-enable Catalyst [releng]
13:34 <dcaro@cloudcumin1001> END (PASS) - Cookbook wmcs.toolforge.component.deploy (exit_code=0) for component components-api [tools]
13:34 <klausman@cumin1003> END (PASS) - Cookbook sre.k8s.pool-depool-node (exit_code=0) depool for host ml-staging2001.codfw.wmnet [production]
13:31 <dpogorzelski@deploy1003> helmfile [ml-staging-codfw] DONE helmfile.d/admin 'sync'. [production]
13:29 <dcaro@cloudcumin1001> START - Cookbook wmcs.toolforge.component.deploy for component components-api [tools]
13:29 <dpogorzelski@deploy1003> helmfile [ml-staging-codfw] START helmfile.d/admin 'sync'. [production]
13:28 <dcaro@cloudcumin1001> END (PASS) - Cookbook wmcs.toolforge.component.deploy (exit_code=0) for component components-api [toolsbeta]
13:28 <fnegri@cumin1003> conftool action : set/pooled=no; selector: name=clouddb1025.eqiad.wmnet [production]
13:26 <btullis@cumin1003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host stat1009.eqiad.wmnet with OS bookworm [production]
13:24 <dcaro@cloudcumin1001> START - Cookbook wmcs.toolforge.component.deploy for component components-api [toolsbeta]
13:24 <klausman@cumin1003> START - Cookbook sre.k8s.pool-depool-node depool for host ml-staging2001.codfw.wmnet [production]
13:24 <klausman@cumin1003> START - Cookbook sre.k8s.reboot-nodes rolling reboot on A:ml-staging-worker [production]
13:23 <klausman@cumin1003> END (PASS) - Cookbook sre.ganeti.reboot-vm (exit_code=0) for VM ml-serve-ctrl1002.eqiad.wmnet [production]
13:21 <btullis@cumin1003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host stat1010.eqiad.wmnet with OS bookworm [production]
13:20 <klausman@cumin1003> START - Cookbook sre.ganeti.reboot-vm for VM ml-serve-ctrl1002.eqiad.wmnet [production]
13:20 <klausman@cumin1003> END (PASS) - Cookbook sre.ganeti.reboot-vm (exit_code=0) for VM ml-serve-ctrl1001.eqiad.wmnet [production]
13:17 <klausman@cumin1003> START - Cookbook sre.ganeti.reboot-vm for VM ml-serve-ctrl1001.eqiad.wmnet [production]
13:16 <klausman@cumin1003> END (PASS) - Cookbook sre.ganeti.reboot-vm (exit_code=0) for VM ml-serve-ctrl2001.codfw.wmnet [production]
13:13 <mszwarc@deploy1003> Finished scap sync-world: Backport for [[gerrit:1327511|UIC: Fix page:page instead of page:other in instrumentation]] (duration: 07m 00s) [production]
13:13 <klausman@cumin1003> START - Cookbook sre.ganeti.reboot-vm for VM ml-serve-ctrl2001.codfw.wmnet [production]
13:12 <klausman@cumin1003> END (PASS) - Cookbook sre.ganeti.reboot-vm (exit_code=0) for VM ml-serve-ctrl2002.codfw.wmnet [production]
13:10 <godog> enable maintain-dumps-nfs - T432583 [paws]
13:09 <mszwarc@deploy1003> mszwarc: Continuing with deployment [production]
13:08 <mszwarc@deploy1003> mszwarc: Backport for [[gerrit:1327511|UIC: Fix page:page instead of page:other in instrumentation]] synced to the testservers (see https://wikitech.wikimedia.org/wiki/Mwdebug). Changes can now be verified there. [production]
13:08 <klausman@cumin1003> START - Cookbook sre.ganeti.reboot-vm for VM ml-serve-ctrl2002.codfw.wmnet [production]