351-400 of 10000 results (46ms)
2026-09-30 ยง
11:23 <cgoubert@deploy1003> helmfile [eqiad] START helmfile.d/services/mw-parsoid: apply [production]
11:22 <cgoubert@deploy1003> helmfile [codfw] DONE helmfile.d/services/mw-jobrunner: apply [production]
11:22 <cgoubert@deploy1003> helmfile [codfw] START helmfile.d/services/mw-jobrunner: apply [production]
11:22 <cgoubert@deploy1003> helmfile [eqiad] DONE helmfile.d/services/mw-jobrunner: apply [production]
11:22 <cgoubert@deploy1003> helmfile [eqiad] START helmfile.d/services/mw-jobrunner: apply [production]
11:07 <cgoubert@deploy1003> Finished scap sync-world: 1330505: mediawiki: Redirect /api/ to /w/rest.php | T433547 (duration: 11m 12s) [production]
11:01 <cgoubert@deploy1003> cgoubert: Continuing with deployment [production]
10:59 <cgoubert@deploy1003> cgoubert: 1330505: mediawiki: Redirect /api/ to /w/rest.php | T433547 synced to the testservers (see https://wikitech.wikimedia.org/wiki/Mwdebug). Changes can now be verified there. [production]
10:56 <cgoubert@deploy1003> Started scap sync-world: 1330505: mediawiki: Redirect /api/ to /w/rest.php | T433547 [production]
10:40 <oblivian@cumin1004> END (PASS) - Cookbook sre.k8s.reboot-nodes (exit_code=0) rolling reboot on D{wikikube-worker1384.eqiad.wmnet} and (A:wikikube-master-eqiad or A:wikikube-worker-eqiad) [production]
10:40 <oblivian@cumin1004> END (PASS) - Cookbook sre.k8s.pool-depool-node (exit_code=0) pool for host wikikube-worker1384.eqiad.wmnet [production]
10:40 <oblivian@cumin1004> START - Cookbook sre.k8s.pool-depool-node pool for host wikikube-worker1384.eqiad.wmnet [production]
10:38 <elukey> `sudo ipvsadm --delete-service --tcp-service 10.2.1.96:443` on lvs201[34] to remove the pki service [production]
10:33 <oblivian@cumin1004> END (PASS) - Cookbook sre.k8s.pool-depool-node (exit_code=0) depool for host wikikube-worker1384.eqiad.wmnet [production]
10:33 <oblivian@cumin1004> START - Cookbook sre.k8s.pool-depool-node depool for host wikikube-worker1384.eqiad.wmnet [production]
10:33 <oblivian@cumin1004> START - Cookbook sre.k8s.reboot-nodes rolling reboot on D{wikikube-worker1384.eqiad.wmnet} and (A:wikikube-master-eqiad or A:wikikube-worker-eqiad) [production]
10:32 <tappof> Grafana upgrade from 13.2.0 to 13.2.3 - T438454 [production]
10:27 <oblivian@cumin1004> END (PASS) - Cookbook sre.hosts.provision (exit_code=0) for host wikikube-worker1384.mgmt.eqiad.wmnet with chassis set policy GRACEFUL_RESTART [production]
10:27 <oblivian@cumin1004> START - Cookbook sre.hosts.provision for host wikikube-worker1384.mgmt.eqiad.wmnet with chassis set policy GRACEFUL_RESTART [production]
10:06 <jelto@cumin1004> END (PASS) - Cookbook sre.loadbalancer.migrate-service-ipip (exit_code=0) for alias: wikikube-worker-eqiad@eqiad [production]
10:06 <jelto@cumin1004> END (PASS) - Cookbook sre.loadbalancer.restart-pybal (exit_code=0) rolling-restart of pybal on (A:lvs-low-traffic-eqiad or A:lvs-secondary-eqiad) and A:bullseye and A:lvs [production]
10:05 <jelto@cumin1004> START - Cookbook sre.loadbalancer.restart-pybal rolling-restart of pybal on (A:lvs-low-traffic-eqiad or A:lvs-secondary-eqiad) and A:bullseye and A:lvs [production]
10:03 <jelto@cumin1004> START - Cookbook sre.loadbalancer.migrate-service-ipip for alias: wikikube-worker-eqiad@eqiad [production]
09:39 <jelto@cumin1004> END (FAIL) - Cookbook sre.loadbalancer.migrate-service-ipip (exit_code=99) for alias: wikikube-worker-eqiad@eqiad [production]
09:38 <jelto@cumin1004> START - Cookbook sre.loadbalancer.migrate-service-ipip for alias: wikikube-worker-eqiad@eqiad [production]
09:33 <jelto@cumin1004> END (PASS) - Cookbook sre.loadbalancer.migrate-service-ipip (exit_code=0) for alias: wikikube-worker-codfw@codfw [production]
09:33 <jelto@cumin1004> END (PASS) - Cookbook sre.loadbalancer.restart-pybal (exit_code=0) rolling-restart of pybal on (A:lvs-low-traffic-codfw or A:lvs-secondary-codfw) and A:bullseye and A:lvs [production]
09:32 <jelto@cumin1004> START - Cookbook sre.loadbalancer.restart-pybal rolling-restart of pybal on (A:lvs-low-traffic-codfw or A:lvs-secondary-codfw) and A:bullseye and A:lvs [production]
09:28 <jelto@cumin1004> START - Cookbook sre.loadbalancer.migrate-service-ipip for alias: wikikube-worker-codfw@codfw [production]
09:08 <btullis@dns1004> END - running authdns-update [production]
09:06 <btullis@dns1004> START - running authdns-update [production]
09:04 <elukey> elukey@cumin1004:~$ sudo cumin 'A:lvs-low-traffic-codfw' 'systemctl restart pybal.service' [production]
09:01 <elukey> elukey@cumin1004:~$ sudo cumin 'A:lvs-secondary-codfw' 'systemctl restart pybal.service' [production]
08:27 <jmm@cumin2003> END (PASS) - Cookbook sre.swift.roll-restart-reboot-swift-thanos-proxies (exit_code=0) rolling restart_daemons on A:thanos-fe-eqiad [production]
08:25 <jmm@cumin2003> START - Cookbook sre.swift.roll-restart-reboot-swift-thanos-proxies rolling restart_daemons on A:thanos-fe-eqiad [production]
08:24 <elukey@cumin1004> conftool action : set/pooled=yes:weight=1; selector: cluster=pki,service=cfssl-multirootca [production]
08:24 <jmm@cumin2003> END (PASS) - Cookbook sre.swift.roll-restart-reboot-swift-thanos-proxies (exit_code=0) rolling restart_daemons on A:thanos-fe-codfw [production]
08:22 <jmm@cumin2003> START - Cookbook sre.swift.roll-restart-reboot-swift-thanos-proxies rolling restart_daemons on A:thanos-fe-codfw [production]
07:49 <filippo@cumin1004> END (ERROR) - Cookbook sre.hosts.reimage (exit_code=97) for host cloudcephosd1056.eqiad.wmnet with OS bookworm [production]
07:46 <elukey@dns1004> END - running authdns-update [production]
07:44 <kevinbazira@deploy1003> helmfile [ml-staging-codfw] Ran 'sync' command on namespace 'experimental' for release 'llm' . [production]
07:44 <kevinbazira@deploy1003> helmfile [ml-staging-codfw] Ran 'sync' command on namespace 'experimental' for release 'main' . [production]
07:44 <elukey@dns1004> START - running authdns-update [production]
07:35 <moritzm> installing pcre2/openssl security updates [production]
07:18 <filippo@cumin1004> START - Cookbook sre.hosts.reimage for host cloudcephosd1056.eqiad.wmnet with OS bookworm [production]
07:18 <filippo@cumin1004> END (ERROR) - Cookbook sre.hosts.reimage (exit_code=97) for host cloudcephosd1056.eqiad.wmnet with OS bookworm [production]
07:15 <dcausse@deploy1003> Finished scap sync-world: Backport for [[gerrit:1346012|Search: optimize morelike queries]] (duration: 11m 27s) [production]
07:13 <filippo@cumin1004> START - Cookbook sre.hosts.reimage for host cloudcephosd1056.eqiad.wmnet with OS bookworm [production]
07:12 <filippo@cumin1004> END (ERROR) - Cookbook sre.hosts.reimage (exit_code=97) for host cloudcephosd1056.eqiad.wmnet with OS bookworm [production]
07:09 <dcausse@deploy1003> dcausse: Continuing with deployment [production]