|
2026-07-30
§
|
| 06:24 |
<root@cumin1003> |
START - Cookbook sre.mysql.pool pool db1177: Maintenance |
[production] |
| 06:17 |
<cwilliams@cumin1003> |
dbctl commit (dc=all): 'Depooling db1177 (T431660)', diff saved to https://phabricator.wikimedia.org/P95741 and previous config saved to /var/cache/conftool/dbconfig/20260730-061736-cwilliams.json |
[production] |
| 06:17 |
<cwilliams@cumin1003> |
DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 1 day, 0:00:00 on db1177.eqiad.wmnet with reason: Maintenance |
[production] |
| 06:17 |
<root@cumin1003> |
END (PASS) - Cookbook sre.mysql.pool (exit_code=0) pool db1172: Maintenance |
[production] |
| 05:44 |
<marostegui@cumin1003> |
DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 1 day, 0:00:00 on db1218.eqiad.wmnet with reason: crashed |
[production] |
| 05:41 |
<marostegui@cumin1003> |
dbctl commit (dc=all): 'Depool db1217 it crashed', diff saved to https://phabricator.wikimedia.org/P95737 and previous config saved to /var/cache/conftool/dbconfig/20260730-054111-marostegui.json |
[production] |
| 05:34 |
<cwilliams@cumin1003> |
dbctl commit (dc=all): 'Repooling after maintenance db1252 (T431660)', diff saved to https://phabricator.wikimedia.org/P95736 and previous config saved to /var/cache/conftool/dbconfig/20260730-053422-cwilliams.json |
[production] |
| 05:30 |
<root@cumin1003> |
START - Cookbook sre.mysql.pool pool db1172: Maintenance |
[production] |
| 05:24 |
<cwilliams@cumin1003> |
dbctl commit (dc=all): 'Repooling after maintenance db1252', diff saved to https://phabricator.wikimedia.org/P95734 and previous config saved to /var/cache/conftool/dbconfig/20260730-052414-cwilliams.json |
[production] |
| 05:23 |
<cwilliams@cumin1003> |
dbctl commit (dc=all): 'Depooling db1172 (T431660)', diff saved to https://phabricator.wikimedia.org/P95733 and previous config saved to /var/cache/conftool/dbconfig/20260730-052354-cwilliams.json |
[production] |
| 05:23 |
<cwilliams@cumin1003> |
DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 1 day, 0:00:00 on db1172.eqiad.wmnet with reason: Maintenance |
[production] |
| 05:23 |
<root@cumin1003> |
END (PASS) - Cookbook sre.mysql.pool (exit_code=0) pool db1167: Maintenance |
[production] |
| 05:14 |
<cwilliams@cumin1003> |
dbctl commit (dc=all): 'Repooling after maintenance db1252', diff saved to https://phabricator.wikimedia.org/P95731 and previous config saved to /var/cache/conftool/dbconfig/20260730-051406-cwilliams.json |
[production] |
| 05:03 |
<cwilliams@cumin1003> |
dbctl commit (dc=all): 'Repooling after maintenance db1252 (T431660)', diff saved to https://phabricator.wikimedia.org/P95729 and previous config saved to /var/cache/conftool/dbconfig/20260730-050358-cwilliams.json |
[production] |
| 04:47 |
<pt1979@cumin2002> |
END (PASS) - Cookbook sre.dns.netbox (exit_code=0) |
[production] |
| 04:47 |
<pt1979@cumin2002> |
END (PASS) - Cookbook sre.puppet.sync-netbox-hiera (exit_code=0) generate netbox hiera data: "Triggered by cookbooks.sre.dns.netbox: Add mr1 ge-0/0/3 ipv4 - pt1979@cumin2002" |
[production] |
| 04:47 |
<pt1979@cumin2002> |
START - Cookbook sre.puppet.sync-netbox-hiera generate netbox hiera data: "Triggered by cookbooks.sre.dns.netbox: Add mr1 ge-0/0/3 ipv4 - pt1979@cumin2002" |
[production] |
| 04:35 |
<root@cumin1003> |
START - Cookbook sre.mysql.pool pool db1167: Maintenance |
[production] |
| 04:29 |
<cwilliams@cumin1003> |
dbctl commit (dc=all): 'Depooling db1167 (T431660)', diff saved to https://phabricator.wikimedia.org/P95726 and previous config saved to /var/cache/conftool/dbconfig/20260730-042923-cwilliams.json |
[production] |
| 04:29 |
<cwilliams@cumin1003> |
DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2 days, 0:00:00 on 6 hosts with reason: Maintenance |
[production] |
| 04:29 |
<cwilliams@cumin1003> |
DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 1 day, 0:00:00 on db1167.eqiad.wmnet with reason: Maintenance |
[production] |
| 04:22 |
<pt1979@cumin2002> |
START - Cookbook sre.dns.netbox |
[production] |
| 04:03 |
<cwilliams@cumin1003> |
dbctl commit (dc=all): 'Depooling db1252 (T431660)', diff saved to https://phabricator.wikimedia.org/P95725 and previous config saved to /var/cache/conftool/dbconfig/20260730-040337-cwilliams.json |
[production] |
| 04:03 |
<cwilliams@cumin1003> |
DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 1 day, 0:00:00 on db1252.eqiad.wmnet with reason: Maintenance |
[production] |
| 02:07 |
<mwpresync@deploy1003> |
Finished scap build-images: Publishing wmf/next image (duration: 06m 37s) |
[production] |
| 02:00 |
<mwpresync@deploy1003> |
Started scap build-images: Publishing wmf/next image |
[production] |
| 01:38 |
<brett@cumin2002> |
END (PASS) - Cookbook sre.dns.admin (exit_code=0) DNS admin: pool eqsin [reason: Switch upgrade maintenance window complete, T433097] |
[production] |
| 01:38 |
<brett@cumin2002> |
START - Cookbook sre.dns.admin DNS admin: pool eqsin [reason: Switch upgrade maintenance window complete, T433097] |
[production] |
|
2026-07-29
§
|
| 23:57 |
<pt1979@cumin1003> |
DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on mr1-eqsin,mr1-eqsin IPv6,mr1-eqsin.oob,mr1-eqsin.oob IPv6 with reason: connection issue |
[production] |
| 22:54 |
<ssastry@deploy1003> |
helmfile [codfw] DONE helmfile.d/services/mw-parsoid: apply |
[production] |
| 22:53 |
<ssastry@deploy1003> |
helmfile [codfw] START helmfile.d/services/mw-parsoid: apply |
[production] |
| 22:53 |
<ssastry@deploy1003> |
helmfile [eqiad] DONE helmfile.d/services/mw-parsoid: apply |
[production] |
| 22:53 |
<ssastry@deploy1003> |
helmfile [eqiad] START helmfile.d/services/mw-parsoid: apply |
[production] |
| 22:25 |
<bking@cumin2003> |
END (PASS) - Cookbook sre.wdqs.data-transfer (exit_code=0) (T430880, restore data on newly-reimaged host) xfer wdqs-all from wdqs2015.codfw.wmnet -> wdqs2022.codfw.wmnet, repooling source-only afterwards |
[production] |
| 22:20 |
<brett@cumin2002> |
END (PASS) - Cookbook sre.dns.admin (exit_code=0) DNS admin: depool eqsin [reason: Switch upgrade maintenance window, T433097] |
[production] |
| 22:20 |
<brett@cumin2002> |
START - Cookbook sre.dns.admin DNS admin: depool eqsin [reason: Switch upgrade maintenance window, T433097] |
[production] |
| 22:01 |
<apine@deploy1003> |
helmfile [eqiad] DONE helmfile.d/services/wikifunctions: apply |
[production] |
| 22:00 |
<apine@deploy1003> |
helmfile [eqiad] START helmfile.d/services/wikifunctions: apply |
[production] |
| 21:59 |
<apine@deploy1003> |
helmfile [codfw] DONE helmfile.d/services/wikifunctions: apply |
[production] |
| 21:58 |
<apine@deploy1003> |
helmfile [codfw] START helmfile.d/services/wikifunctions: apply |
[production] |
| 21:58 |
<apine@deploy1003> |
helmfile [staging] DONE helmfile.d/services/wikifunctions: apply |
[production] |
| 21:58 |
<apine@deploy1003> |
helmfile [staging] START helmfile.d/services/wikifunctions: apply |
[production] |
| 21:32 |
<vriley@cumin1003> |
END (FAIL) - Cookbook sre.hosts.reimage (exit_code=99) for host cloudvirt1048.eqiad.wmnet with OS trixie |
[production] |
| 21:25 |
<pt1979@cumin2002> |
END (ERROR) - Cookbook sre.hosts.reimage (exit_code=97) for host ms-be2097.codfw.wmnet with OS bullseye |
[production] |
| 21:16 |
<zabe@deploy1003> |
mwscript-k8s job started: extensions/Translate/scripts/moveTranslatableBundle.php --wiki=metawiki 'Mental Health Resource Center' 'Safety Resource Center/Mental Health' Zabe --reason 'per request [[:phab:T433118|T433118]]' |
[production] |
| 21:12 |
<bking@cumin2003> |
START - Cookbook sre.wdqs.data-transfer (T430880, restore data on newly-reimaged host) xfer wdqs-all from wdqs2015.codfw.wmnet -> wdqs2022.codfw.wmnet, repooling source-only afterwards |
[production] |
| 21:12 |
<aaron@deploy1003> |
Finished scap sync-world: Backport for [[gerrit:1319175|Make the "wikibase-rest/v1" external module published (T422405)]] (duration: 12m 53s) |
[production] |
| 21:12 |
<bking@cumin2003> |
conftool action : set/pooled=yes; selector: name=wdqs2015\.codfw\.wmnet,dc=codfw,cluster=wdqs\-main,service=wdqs\-main |
[production] |
| 21:10 |
<root@cumin1003> |
END (PASS) - Cookbook sre.mysql.pool (exit_code=0) pool db1253: Maintenance |
[production] |
| 21:08 |
<aaron@deploy1003> |
aaron: Continuing with deployment |
[production] |