151-200 of 10000 results (49ms)
2026-09-21 ยง
11:35 <brouberol@deploy1003> helmfile [dse-k8s-eqiad] START helmfile.d/dse-k8s-services/growthbook: apply [production]
11:35 <brouberol@deploy1003> helmfile [dse-k8s-eqiad] DONE helmfile.d/dse-k8s-services/growthboo-next: apply [production]
11:34 <jclark@cumin1004> START - Cookbook sre.hosts.reimage for host ml-serve1016.eqiad.wmnet with OS trixie [production]
11:34 <brouberol@deploy1003> helmfile [dse-k8s-eqiad] START helmfile.d/dse-k8s-services/growthbook-next: apply [production]
11:34 <brouberol@deploy1003> helmfile [dse-k8s-eqiad] DONE helmfile.d/dse-k8s-services/superset: apply [production]
11:33 <brouberol@deploy1003> helmfile [dse-k8s-eqiad] START helmfile.d/dse-k8s-services/superset: apply [production]
11:33 <brouberol@deploy1003> helmfile [dse-k8s-eqiad] DONE helmfile.d/dse-k8s-services/superset-next: apply [production]
11:33 <brouberol@deploy1003> helmfile [dse-k8s-eqiad] START helmfile.d/dse-k8s-services/superset-next: apply [production]
11:33 <brouberol@deploy1003> helmfile [dse-k8s-eqiad] DONE helmfile.d/dse-k8s-services/spark-history: apply [production]
11:32 <brouberol@deploy1003> helmfile [dse-k8s-eqiad] START helmfile.d/dse-k8s-services/spark-history: apply [production]
11:32 <brouberol@deploy1003> helmfile [dse-k8s-eqiad] DONE helmfile.d/dse-k8s-services/spark-history: apply [production]
11:31 <jclark@cumin1004> END (FAIL) - Cookbook sre.hosts.provision (exit_code=99) for host ml-serve1016.mgmt.eqiad.wmnet with chassis set policy FORCE_RESTART [production]
11:31 <jclark@cumin1004> START - Cookbook sre.hosts.provision for host ml-serve1016.mgmt.eqiad.wmnet with chassis set policy FORCE_RESTART [production]
11:31 <brouberol@deploy1003> helmfile [dse-k8s-eqiad] START helmfile.d/dse-k8s-services/spark-history: apply [production]
11:13 <btullis@cumin1004> START - Cookbook sre.hosts.reboot-single for host dse-k8s-worker2004.codfw.wmnet [production]
11:13 <btullis@cumin1004> END (PASS) - Cookbook sre.hosts.reboot-single (exit_code=0) for host dse-k8s-worker2003.codfw.wmnet [production]
11:04 <urbanecm@deploy1003> mwscript-k8s job started: extensions/Translate/scripts/moveTranslatableBundle.php --wiki mediawikiwiki 'Wikimedia Apps/Team/Android/Customizable Donation Reminder Experiment' 'Wikimedia Apps/Team/Customizable Donation Reminder/Android' 'Martin Urbanec' --reason 'per request [[:phab:T438704|T438704]]' [production]
10:59 <btullis@cumin1004> START - Cookbook sre.hosts.reboot-single for host dse-k8s-worker2003.codfw.wmnet [production]
10:54 <btullis@cumin1004> END (PASS) - Cookbook sre.hosts.reboot-single (exit_code=0) for host dse-k8s-worker2002.codfw.wmnet [production]
10:50 <urbanecm@deploy1003> mwscript-k8s job started: extensions/Translate/scripts/moveTranslatableBundle.php --wiki mediawikiwiki 'Wikimedia Apps/Team/Android/Customizable Donation Reminder Experiment' 'Wikimedia Apps/Team/Customizable Donation Reminder/Android' Zabe --reason 'per request [[:phab:T438704|T438704]]' [production]
10:38 <zabe> create wbc_entity_usage table in x1 for all wikidata client wikis # T438499 [production]
10:36 <btullis@cumin1004> START - Cookbook sre.hosts.reboot-single for host dse-k8s-worker2002.codfw.wmnet [production]
10:36 <btullis@cumin1004> END (PASS) - Cookbook sre.hosts.reboot-single (exit_code=0) for host dse-k8s-worker2001.codfw.wmnet [production]
10:21 <zabe@deploy1003> mwscript-k8s job started: extensions/Translate/scripts/moveTranslatableBundle.php --wiki mediawikiwiki 'Wikimedia Apps/Team/Android/Customizable Donation Reminder Experiment' 'Wikimedia Apps/Team/Customizable Donation Reminder/Android' Zabe --reason 'per request [[:phab:T438704|T438704]]' [production]
10:21 <btullis@cumin1004> START - Cookbook sre.hosts.reboot-single for host dse-k8s-worker2001.codfw.wmnet [production]
10:21 <zabe@deploy1003> mwscript-k8s job started: extensions/Translate/scripts/moveTranslatableBundle.php --wiki mediawikiwiki 'Wikimedia Apps/Team/Android/Customizable Donation Reminder Experiment' 'Wikimedia Apps/Team/Customizable Donation Reminder/Android' Zabe --reason 'per request [[:phab:T438704|T438704]]' [production]
10:19 <zabe@deploy1003> mwscript-k8s job started: extensions/Translate/scripts/moveTranslatableBundle.php --wiki metawiki 'Wikimedia Apps/Team/Android/Customizable Donation Reminder Experiment' 'Wikimedia Apps/Team/Customizable Donation Reminder/Android' Zabe --reason 'per request [[:phab:T438704|T438704]]' [production]
10:17 <jelto@cumin1004> END (PASS) - Cookbook sre.loadbalancer.migrate-service-ipip (exit_code=0) for alias: wikikube-worker-eqiad@eqiad [production]
10:17 <jelto@cumin1004> END (PASS) - Cookbook sre.loadbalancer.restart-pybal (exit_code=0) rolling-restart of pybal on (A:lvs-low-traffic-eqiad or A:lvs-secondary-eqiad) and A:bullseye and A:lvs [production]
10:16 <jelto@cumin1004> START - Cookbook sre.loadbalancer.restart-pybal rolling-restart of pybal on (A:lvs-low-traffic-eqiad or A:lvs-secondary-eqiad) and A:bullseye and A:lvs [production]
10:12 <jmm@deploy1003> helmfile [eqiad] DONE helmfile.d/services/proton: apply [production]
10:11 <jelto@cumin1004> START - Cookbook sre.loadbalancer.migrate-service-ipip for alias: wikikube-worker-eqiad@eqiad [production]
10:09 <jmm@deploy1003> helmfile [eqiad] START helmfile.d/services/proton: apply [production]
10:04 <jmm@deploy1003> helmfile [codfw] DONE helmfile.d/services/proton: apply [production]
10:02 <jmm@deploy1003> helmfile [codfw] START helmfile.d/services/proton: apply [production]
10:01 <jmm@deploy1003> helmfile [staging] DONE helmfile.d/services/proton: apply [production]
10:00 <jmm@deploy1003> helmfile [staging] START helmfile.d/services/proton: apply [production]
10:00 <jmm@deploy1003> helmfile [staging] DONE helmfile.d/services/proton: apply [production]
09:59 <filippo@cumin1004> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host cloudvirt1063.eqiad.wmnet with OS trixie [production]
09:59 <jmm@deploy1003> helmfile [staging] START helmfile.d/services/proton: apply [production]
09:56 <klausman@deploy1003> helmfile [dse-k8s-eqiad] DONE helmfile.d/dse-k8s-services/liftwing-studio: apply [production]
09:55 <jelto@cumin1004> END (PASS) - Cookbook sre.loadbalancer.restart-pybal (exit_code=0) rolling-restart of pybal on (A:lvs-low-traffic-codfw or A:lvs-secondary-codfw) and A:bullseye and A:lvs [production]
09:55 <jelto@cumin1004> END (PASS) - Cookbook sre.loadbalancer.migrate-service-ipip (exit_code=0) for alias: wikikube-worker-codfw@codfw [production]
09:54 <jelto@cumin1004> START - Cookbook sre.loadbalancer.restart-pybal rolling-restart of pybal on (A:lvs-low-traffic-codfw or A:lvs-secondary-codfw) and A:bullseye and A:lvs [production]
09:54 <klausman@deploy1003> helmfile [dse-k8s-eqiad] START helmfile.d/dse-k8s-services/liftwing-studio: apply [production]
09:50 <jelto@cumin1004> START - Cookbook sre.loadbalancer.migrate-service-ipip for alias: wikikube-worker-codfw@codfw [production]
09:35 <moritzm> installing chromium security updates [production]
09:22 <tappof> bump space for prometheus k8s-dse in eqiad [production]
09:11 <ihurbain@deploy1003> helmfile [codfw] DONE helmfile.d/services/mw-parsoid: apply [production]
09:07 <filippo@cumin1004> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on cloudvirt1063.eqiad.wmnet with reason: host reimage [production]