|
2026-09-30
ยง
|
| 11:23 |
<cgoubert@deploy1003> |
helmfile [eqiad] START helmfile.d/services/mw-parsoid: apply |
[production] |
| 11:22 |
<cgoubert@deploy1003> |
helmfile [codfw] DONE helmfile.d/services/mw-jobrunner: apply |
[production] |
| 11:22 |
<cgoubert@deploy1003> |
helmfile [codfw] START helmfile.d/services/mw-jobrunner: apply |
[production] |
| 11:22 |
<cgoubert@deploy1003> |
helmfile [eqiad] DONE helmfile.d/services/mw-jobrunner: apply |
[production] |
| 11:22 |
<cgoubert@deploy1003> |
helmfile [eqiad] START helmfile.d/services/mw-jobrunner: apply |
[production] |
| 11:07 |
<cgoubert@deploy1003> |
Finished scap sync-world: 1330505: mediawiki: Redirect /api/ to /w/rest.php | T433547 (duration: 11m 12s) |
[production] |
| 11:01 |
<cgoubert@deploy1003> |
cgoubert: Continuing with deployment |
[production] |
| 10:59 |
<cgoubert@deploy1003> |
cgoubert: 1330505: mediawiki: Redirect /api/ to /w/rest.php | T433547 synced to the testservers (see https://wikitech.wikimedia.org/wiki/Mwdebug). Changes can now be verified there. |
[production] |
| 10:56 |
<cgoubert@deploy1003> |
Started scap sync-world: 1330505: mediawiki: Redirect /api/ to /w/rest.php | T433547 |
[production] |
| 10:40 |
<oblivian@cumin1004> |
END (PASS) - Cookbook sre.k8s.reboot-nodes (exit_code=0) rolling reboot on D{wikikube-worker1384.eqiad.wmnet} and (A:wikikube-master-eqiad or A:wikikube-worker-eqiad) |
[production] |
| 10:40 |
<oblivian@cumin1004> |
END (PASS) - Cookbook sre.k8s.pool-depool-node (exit_code=0) pool for host wikikube-worker1384.eqiad.wmnet |
[production] |
| 10:40 |
<oblivian@cumin1004> |
START - Cookbook sre.k8s.pool-depool-node pool for host wikikube-worker1384.eqiad.wmnet |
[production] |
| 10:38 |
<elukey> |
`sudo ipvsadm --delete-service --tcp-service 10.2.1.96:443` on lvs201[34] to remove the pki service |
[production] |
| 10:33 |
<oblivian@cumin1004> |
END (PASS) - Cookbook sre.k8s.pool-depool-node (exit_code=0) depool for host wikikube-worker1384.eqiad.wmnet |
[production] |
| 10:33 |
<oblivian@cumin1004> |
START - Cookbook sre.k8s.pool-depool-node depool for host wikikube-worker1384.eqiad.wmnet |
[production] |
| 10:33 |
<oblivian@cumin1004> |
START - Cookbook sre.k8s.reboot-nodes rolling reboot on D{wikikube-worker1384.eqiad.wmnet} and (A:wikikube-master-eqiad or A:wikikube-worker-eqiad) |
[production] |
| 10:32 |
<tappof> |
Grafana upgrade from 13.2.0 to 13.2.3 - T438454 |
[production] |
| 10:27 |
<oblivian@cumin1004> |
END (PASS) - Cookbook sre.hosts.provision (exit_code=0) for host wikikube-worker1384.mgmt.eqiad.wmnet with chassis set policy GRACEFUL_RESTART |
[production] |
| 10:27 |
<oblivian@cumin1004> |
START - Cookbook sre.hosts.provision for host wikikube-worker1384.mgmt.eqiad.wmnet with chassis set policy GRACEFUL_RESTART |
[production] |
| 10:06 |
<jelto@cumin1004> |
END (PASS) - Cookbook sre.loadbalancer.migrate-service-ipip (exit_code=0) for alias: wikikube-worker-eqiad@eqiad |
[production] |
| 10:06 |
<jelto@cumin1004> |
END (PASS) - Cookbook sre.loadbalancer.restart-pybal (exit_code=0) rolling-restart of pybal on (A:lvs-low-traffic-eqiad or A:lvs-secondary-eqiad) and A:bullseye and A:lvs |
[production] |
| 10:05 |
<jelto@cumin1004> |
START - Cookbook sre.loadbalancer.restart-pybal rolling-restart of pybal on (A:lvs-low-traffic-eqiad or A:lvs-secondary-eqiad) and A:bullseye and A:lvs |
[production] |
| 10:03 |
<jelto@cumin1004> |
START - Cookbook sre.loadbalancer.migrate-service-ipip for alias: wikikube-worker-eqiad@eqiad |
[production] |
| 09:39 |
<jelto@cumin1004> |
END (FAIL) - Cookbook sre.loadbalancer.migrate-service-ipip (exit_code=99) for alias: wikikube-worker-eqiad@eqiad |
[production] |
| 09:38 |
<jelto@cumin1004> |
START - Cookbook sre.loadbalancer.migrate-service-ipip for alias: wikikube-worker-eqiad@eqiad |
[production] |
| 09:33 |
<jelto@cumin1004> |
END (PASS) - Cookbook sre.loadbalancer.migrate-service-ipip (exit_code=0) for alias: wikikube-worker-codfw@codfw |
[production] |
| 09:33 |
<jelto@cumin1004> |
END (PASS) - Cookbook sre.loadbalancer.restart-pybal (exit_code=0) rolling-restart of pybal on (A:lvs-low-traffic-codfw or A:lvs-secondary-codfw) and A:bullseye and A:lvs |
[production] |
| 09:32 |
<jelto@cumin1004> |
START - Cookbook sre.loadbalancer.restart-pybal rolling-restart of pybal on (A:lvs-low-traffic-codfw or A:lvs-secondary-codfw) and A:bullseye and A:lvs |
[production] |
| 09:28 |
<jelto@cumin1004> |
START - Cookbook sre.loadbalancer.migrate-service-ipip for alias: wikikube-worker-codfw@codfw |
[production] |
| 09:08 |
<btullis@dns1004> |
END - running authdns-update |
[production] |
| 09:06 |
<btullis@dns1004> |
START - running authdns-update |
[production] |
| 09:04 |
<elukey> |
elukey@cumin1004:~$ sudo cumin 'A:lvs-low-traffic-codfw' 'systemctl restart pybal.service' |
[production] |
| 09:01 |
<elukey> |
elukey@cumin1004:~$ sudo cumin 'A:lvs-secondary-codfw' 'systemctl restart pybal.service' |
[production] |
| 08:27 |
<jmm@cumin2003> |
END (PASS) - Cookbook sre.swift.roll-restart-reboot-swift-thanos-proxies (exit_code=0) rolling restart_daemons on A:thanos-fe-eqiad |
[production] |
| 08:25 |
<jmm@cumin2003> |
START - Cookbook sre.swift.roll-restart-reboot-swift-thanos-proxies rolling restart_daemons on A:thanos-fe-eqiad |
[production] |
| 08:24 |
<elukey@cumin1004> |
conftool action : set/pooled=yes:weight=1; selector: cluster=pki,service=cfssl-multirootca |
[production] |
| 08:24 |
<jmm@cumin2003> |
END (PASS) - Cookbook sre.swift.roll-restart-reboot-swift-thanos-proxies (exit_code=0) rolling restart_daemons on A:thanos-fe-codfw |
[production] |
| 08:22 |
<jmm@cumin2003> |
START - Cookbook sre.swift.roll-restart-reboot-swift-thanos-proxies rolling restart_daemons on A:thanos-fe-codfw |
[production] |
| 07:49 |
<filippo@cumin1004> |
END (ERROR) - Cookbook sre.hosts.reimage (exit_code=97) for host cloudcephosd1056.eqiad.wmnet with OS bookworm |
[production] |
| 07:46 |
<elukey@dns1004> |
END - running authdns-update |
[production] |
| 07:44 |
<kevinbazira@deploy1003> |
helmfile [ml-staging-codfw] Ran 'sync' command on namespace 'experimental' for release 'llm' . |
[production] |
| 07:44 |
<kevinbazira@deploy1003> |
helmfile [ml-staging-codfw] Ran 'sync' command on namespace 'experimental' for release 'main' . |
[production] |
| 07:44 |
<elukey@dns1004> |
START - running authdns-update |
[production] |
| 07:35 |
<moritzm> |
installing pcre2/openssl security updates |
[production] |
| 07:18 |
<filippo@cumin1004> |
START - Cookbook sre.hosts.reimage for host cloudcephosd1056.eqiad.wmnet with OS bookworm |
[production] |
| 07:18 |
<filippo@cumin1004> |
END (ERROR) - Cookbook sre.hosts.reimage (exit_code=97) for host cloudcephosd1056.eqiad.wmnet with OS bookworm |
[production] |
| 07:15 |
<dcausse@deploy1003> |
Finished scap sync-world: Backport for [[gerrit:1346012|Search: optimize morelike queries]] (duration: 11m 27s) |
[production] |
| 07:13 |
<filippo@cumin1004> |
START - Cookbook sre.hosts.reimage for host cloudcephosd1056.eqiad.wmnet with OS bookworm |
[production] |
| 07:12 |
<filippo@cumin1004> |
END (ERROR) - Cookbook sre.hosts.reimage (exit_code=97) for host cloudcephosd1056.eqiad.wmnet with OS bookworm |
[production] |
| 07:09 |
<dcausse@deploy1003> |
dcausse: Continuing with deployment |
[production] |