|
2026-07-30
§
|
| 04:47 |
<pt1979@cumin2002> |
END (PASS) - Cookbook sre.dns.netbox (exit_code=0) |
[production] |
| 04:47 |
<pt1979@cumin2002> |
END (PASS) - Cookbook sre.puppet.sync-netbox-hiera (exit_code=0) generate netbox hiera data: "Triggered by cookbooks.sre.dns.netbox: Add mr1 ge-0/0/3 ipv4 - pt1979@cumin2002" |
[production] |
| 04:47 |
<pt1979@cumin2002> |
START - Cookbook sre.puppet.sync-netbox-hiera generate netbox hiera data: "Triggered by cookbooks.sre.dns.netbox: Add mr1 ge-0/0/3 ipv4 - pt1979@cumin2002" |
[production] |
| 04:35 |
<root@cumin1003> |
START - Cookbook sre.mysql.pool pool db1167: Maintenance |
[production] |
| 04:29 |
<cwilliams@cumin1003> |
dbctl commit (dc=all): 'Depooling db1167 (T431660)', diff saved to https://phabricator.wikimedia.org/P95726 and previous config saved to /var/cache/conftool/dbconfig/20260730-042923-cwilliams.json |
[production] |
| 04:29 |
<cwilliams@cumin1003> |
DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2 days, 0:00:00 on 6 hosts with reason: Maintenance |
[production] |
| 04:29 |
<cwilliams@cumin1003> |
DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 1 day, 0:00:00 on db1167.eqiad.wmnet with reason: Maintenance |
[production] |
| 04:22 |
<pt1979@cumin2002> |
START - Cookbook sre.dns.netbox |
[production] |
| 04:03 |
<cwilliams@cumin1003> |
dbctl commit (dc=all): 'Depooling db1252 (T431660)', diff saved to https://phabricator.wikimedia.org/P95725 and previous config saved to /var/cache/conftool/dbconfig/20260730-040337-cwilliams.json |
[production] |
| 04:03 |
<cwilliams@cumin1003> |
DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 1 day, 0:00:00 on db1252.eqiad.wmnet with reason: Maintenance |
[production] |
| 02:07 |
<mwpresync@deploy1003> |
Finished scap build-images: Publishing wmf/next image (duration: 06m 37s) |
[production] |
| 02:00 |
<mwpresync@deploy1003> |
Started scap build-images: Publishing wmf/next image |
[production] |
| 01:38 |
<brett@cumin2002> |
END (PASS) - Cookbook sre.dns.admin (exit_code=0) DNS admin: pool eqsin [reason: Switch upgrade maintenance window complete, T433097] |
[production] |
| 01:38 |
<brett@cumin2002> |
START - Cookbook sre.dns.admin DNS admin: pool eqsin [reason: Switch upgrade maintenance window complete, T433097] |
[production] |
|
2026-07-29
§
|
| 23:57 |
<pt1979@cumin1003> |
DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on mr1-eqsin,mr1-eqsin IPv6,mr1-eqsin.oob,mr1-eqsin.oob IPv6 with reason: connection issue |
[production] |
| 22:54 |
<ssastry@deploy1003> |
helmfile [codfw] DONE helmfile.d/services/mw-parsoid: apply |
[production] |
| 22:53 |
<ssastry@deploy1003> |
helmfile [codfw] START helmfile.d/services/mw-parsoid: apply |
[production] |
| 22:53 |
<ssastry@deploy1003> |
helmfile [eqiad] DONE helmfile.d/services/mw-parsoid: apply |
[production] |
| 22:53 |
<ssastry@deploy1003> |
helmfile [eqiad] START helmfile.d/services/mw-parsoid: apply |
[production] |
| 22:25 |
<bking@cumin2003> |
END (PASS) - Cookbook sre.wdqs.data-transfer (exit_code=0) (T430880, restore data on newly-reimaged host) xfer wdqs-all from wdqs2015.codfw.wmnet -> wdqs2022.codfw.wmnet, repooling source-only afterwards |
[production] |
| 22:20 |
<brett@cumin2002> |
END (PASS) - Cookbook sre.dns.admin (exit_code=0) DNS admin: depool eqsin [reason: Switch upgrade maintenance window, T433097] |
[production] |
| 22:20 |
<brett@cumin2002> |
START - Cookbook sre.dns.admin DNS admin: depool eqsin [reason: Switch upgrade maintenance window, T433097] |
[production] |
| 22:01 |
<apine@deploy1003> |
helmfile [eqiad] DONE helmfile.d/services/wikifunctions: apply |
[production] |
| 22:00 |
<apine@deploy1003> |
helmfile [eqiad] START helmfile.d/services/wikifunctions: apply |
[production] |
| 21:59 |
<apine@deploy1003> |
helmfile [codfw] DONE helmfile.d/services/wikifunctions: apply |
[production] |
| 21:58 |
<apine@deploy1003> |
helmfile [codfw] START helmfile.d/services/wikifunctions: apply |
[production] |
| 21:58 |
<apine@deploy1003> |
helmfile [staging] DONE helmfile.d/services/wikifunctions: apply |
[production] |
| 21:58 |
<apine@deploy1003> |
helmfile [staging] START helmfile.d/services/wikifunctions: apply |
[production] |
| 21:32 |
<vriley@cumin1003> |
END (FAIL) - Cookbook sre.hosts.reimage (exit_code=99) for host cloudvirt1048.eqiad.wmnet with OS trixie |
[production] |
| 21:25 |
<pt1979@cumin2002> |
END (ERROR) - Cookbook sre.hosts.reimage (exit_code=97) for host ms-be2097.codfw.wmnet with OS bullseye |
[production] |
| 21:16 |
<zabe@deploy1003> |
mwscript-k8s job started: extensions/Translate/scripts/moveTranslatableBundle.php --wiki=metawiki 'Mental Health Resource Center' 'Safety Resource Center/Mental Health' Zabe --reason 'per request [[:phab:T433118|T433118]]' |
[production] |
| 21:12 |
<bking@cumin2003> |
START - Cookbook sre.wdqs.data-transfer (T430880, restore data on newly-reimaged host) xfer wdqs-all from wdqs2015.codfw.wmnet -> wdqs2022.codfw.wmnet, repooling source-only afterwards |
[production] |
| 21:12 |
<aaron@deploy1003> |
Finished scap sync-world: Backport for [[gerrit:1319175|Make the "wikibase-rest/v1" external module published (T422405)]] (duration: 12m 53s) |
[production] |
| 21:12 |
<bking@cumin2003> |
conftool action : set/pooled=yes; selector: name=wdqs2015\.codfw\.wmnet,dc=codfw,cluster=wdqs\-main,service=wdqs\-main |
[production] |
| 21:10 |
<root@cumin1003> |
END (PASS) - Cookbook sre.mysql.pool (exit_code=0) pool db1253: Maintenance |
[production] |
| 21:08 |
<aaron@deploy1003> |
aaron: Continuing with deployment |
[production] |
| 21:01 |
<aaron@deploy1003> |
aaron: Backport for [[gerrit:1319175|Make the "wikibase-rest/v1" external module published (T422405)]] synced to the testservers (see https://wikitech.wikimedia.org/wiki/Mwdebug). Changes can now be verified there. |
[production] |
| 20:59 |
<aaron@deploy1003> |
Started scap sync-world: Backport for [[gerrit:1319175|Make the "wikibase-rest/v1" external module published (T422405)]] |
[production] |
| 20:52 |
<aaron@deploy1003> |
Finished scap sync-world: Backport for [[gerrit:1312694|Use the openapi.json endpoint for the wikibase-rest/v1 REST module (T422405)]] (duration: 21m 57s) |
[production] |
| 20:48 |
<aaron@deploy1003> |
aaron: Continuing with deployment |
[production] |
| 20:32 |
<aaron@deploy1003> |
aaron: Backport for [[gerrit:1312694|Use the openapi.json endpoint for the wikibase-rest/v1 REST module (T422405)]] synced to the testservers (see https://wikitech.wikimedia.org/wiki/Mwdebug). Changes can now be verified there. |
[production] |
| 20:30 |
<aaron@deploy1003> |
Started scap sync-world: Backport for [[gerrit:1312694|Use the openapi.json endpoint for the wikibase-rest/v1 REST module (T422405)]] |
[production] |
| 20:24 |
<root@cumin1003> |
START - Cookbook sre.mysql.pool pool db1253: Maintenance |
[production] |
| 20:19 |
<aaron@deploy1003> |
Finished scap sync-world: Backport for [[gerrit:1306988|REST: use RestExternalModules config variable (T433314 T428375)]] (duration: 08m 07s) |
[production] |
| 20:18 |
<cwilliams@cumin1003> |
dbctl commit (dc=all): 'Depooling db1253 (T431660)', diff saved to https://phabricator.wikimedia.org/P95719 and previous config saved to /var/cache/conftool/dbconfig/20260729-201810-cwilliams.json |
[production] |
| 20:18 |
<cwilliams@cumin1003> |
DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 1 day, 0:00:00 on db1253.eqiad.wmnet with reason: Maintenance |
[production] |
| 20:17 |
<root@cumin1003> |
END (PASS) - Cookbook sre.mysql.pool (exit_code=0) pool db1231: Maintenance |
[production] |
| 20:15 |
<aaron@deploy1003> |
bpirkle, aaron: Continuing with deployment |
[production] |
| 20:13 |
<aaron@deploy1003> |
bpirkle, aaron: Backport for [[gerrit:1306988|REST: use RestExternalModules config variable (T433314 T428375)]] synced to the testservers (see https://wikitech.wikimedia.org/wiki/Mwdebug). Changes can now be verified there. |
[production] |
| 20:12 |
<vriley@cumin1003> |
START - Cookbook sre.hosts.reimage for host cloudvirt1048.eqiad.wmnet with OS trixie |
[production] |