|
2026-09-03
ยง
|
| 11:58 |
<kart_> |
cxserver: Use urldownloader LVS endpoint (T429175) |
[production] |
| 11:57 |
<kartik@deploy1003> |
helmfile [eqiad] DONE helmfile.d/services/cxserver: apply |
[production] |
| 11:56 |
<kartik@deploy1003> |
helmfile [eqiad] START helmfile.d/services/cxserver: apply |
[production] |
| 11:56 |
<slyngshede@cumin1003> |
START - Cookbook sre.cdn.roll-restart-purged rolling restart_daemons on A:cp-upload_magru |
[production] |
| 11:55 |
<kartik@deploy1003> |
helmfile [codfw] DONE helmfile.d/services/cxserver: apply |
[production] |
| 11:55 |
<moritzm> |
installing rsync security updates |
[production] |
| 11:55 |
<kartik@deploy1003> |
helmfile [codfw] START helmfile.d/services/cxserver: apply |
[production] |
| 11:54 |
<mvernon@cumin2003> |
END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on aqs1020.eqiad.wmnet with reason: host reimage |
[production] |
| 11:52 |
<kartik@deploy1003> |
helmfile [staging] DONE helmfile.d/services/cxserver: apply |
[production] |
| 11:51 |
<kartik@deploy1003> |
helmfile [staging] START helmfile.d/services/cxserver: apply |
[production] |
| 11:47 |
<mvernon@cumin2003> |
START - Cookbook sre.hosts.downtime for 2:00:00 on aqs1020.eqiad.wmnet with reason: host reimage |
[production] |
| 11:46 |
<marostegui@cumin1003> |
START - Cookbook sre.mysql.pool pool db1282: Pooling db1282 into s6 |
[production] |
| 11:45 |
<marostegui@cumin1003> |
dbctl commit (dc=all): 'Add db1282 to dbctl T407942', diff saved to https://phabricator.wikimedia.org/P96328 and previous config saved to /var/cache/conftool/dbconfig/20260903-114526-marostegui.json |
[production] |
| 11:43 |
<slyngshede@cumin1003> |
END (PASS) - Cookbook sre.cdn.roll-restart-purged (exit_code=0) rolling restart_daemons on A:cp-upload_eqiad |
[production] |
| 11:35 |
<slyngshede@cumin1003> |
START - Cookbook sre.cdn.roll-restart-purged rolling restart_daemons on A:cp-upload_eqiad |
[production] |
| 11:32 |
<mvernon@cumin2003> |
START - Cookbook sre.hosts.reimage for host aqs1020.eqiad.wmnet with OS bookworm |
[production] |
| 11:26 |
<slyngshede@cumin1003> |
END (PASS) - Cookbook sre.cdn.roll-restart-purged (exit_code=0) rolling restart_daemons on A:cp-upload_eqsin |
[production] |
| 11:26 |
<hashar> |
Deleted coverage report for SimilarEditors ( /srv/doc/cover-extensions/SimilarEditors ), extension is being archived # T436880 |
[releng] |
| 11:24 |
<cgoubert@deploy1003> |
Finished scap sync-world: mediawiki: enable forward of fatal metrics to statsd exporter (duration: 12m 01s) |
[production] |
| 11:24 |
<marostegui@cumin1003> |
START - Cookbook sre.mysql.pool pool db2207: db2207 repool |
[production] |
| 11:22 |
<cgoubert@deploy1003> |
cgoubert: Continuing with deployment |
[production] |
| 11:18 |
<slyngshede@cumin1003> |
START - Cookbook sre.cdn.roll-restart-purged rolling restart_daemons on A:cp-upload_eqsin |
[production] |
| 11:17 |
<slyngshede@cumin1003> |
END (PASS) - Cookbook sre.cdn.roll-restart-purged (exit_code=0) rolling restart_daemons on A:cp-upload_esams |
[production] |
| 11:15 |
<cgoubert@deploy1003> |
cgoubert: mediawiki: enable forward of fatal metrics to statsd exporter synced to the testservers (see https://wikitech.wikimedia.org/wiki/Mwdebug). Changes can now be verified there. |
[production] |
| 11:14 |
<cgoubert@deploy1003> |
Started scap sync-world: mediawiki: enable forward of fatal metrics to statsd exporter |
[production] |
| 11:10 |
<slyngshede@cumin1003> |
START - Cookbook sre.cdn.roll-restart-purged rolling restart_daemons on A:cp-upload_esams |
[production] |
| 11:09 |
<slyngshede@cumin1003> |
END (PASS) - Cookbook sre.cdn.roll-restart-purged (exit_code=0) rolling restart_daemons on A:cp-upload_drmrs |
[production] |
| 11:01 |
<mvernon@cumin2003> |
END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host aqs1019.eqiad.wmnet with OS bookworm |
[production] |
| 11:01 |
<slyngshede@cumin1003> |
START - Cookbook sre.cdn.roll-restart-purged rolling restart_daemons on A:cp-upload_drmrs |
[production] |
| 10:59 |
<slyngshede@cumin1003> |
END (PASS) - Cookbook sre.cdn.roll-restart-purged (exit_code=0) rolling restart_daemons on A:cp-upload_codfw |
[production] |
| 10:52 |
<slyngshede@cumin1003> |
START - Cookbook sre.cdn.roll-restart-purged rolling restart_daemons on A:cp-upload_codfw |
[production] |
| 10:43 |
<mvernon@cumin2003> |
END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on aqs1019.eqiad.wmnet with reason: host reimage |
[production] |
| 10:41 |
<gkyziridis@deploy1003> |
helmfile [ml-serve-eqiad] Ran 'sync' command on namespace 'experimental' for release 'main' . |
[production] |
| 10:40 |
<elukey@cumin1003> |
END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host cassandra-dev2001.codfw.wmnet with OS bookworm |
[production] |
| 10:39 |
<mvernon@cumin2003> |
START - Cookbook sre.hosts.downtime for 2:00:00 on aqs1019.eqiad.wmnet with reason: host reimage |
[production] |
| 10:29 |
<ayounsi@cumin1003> |
END (PASS) - Cookbook sre.ganeti.makevm (exit_code=0) for new host netflow2005.codfw.wmnet |
[production] |
| 10:29 |
<ayounsi@cumin1003> |
END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host netflow2005.codfw.wmnet with OS trixie |
[production] |
| 10:20 |
<elukey@cumin1003> |
END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on cassandra-dev2001.codfw.wmnet with reason: host reimage |
[production] |
| 10:17 |
<btullis@deploy1003> |
helmfile [codfw] DONE helmfile.d/services/eventgate-analytics: sync |
[production] |
| 10:17 |
<btullis@deploy1003> |
helmfile [codfw] START helmfile.d/services/eventgate-analytics: sync |
[production] |
| 10:16 |
<elukey@cumin1003> |
START - Cookbook sre.hosts.downtime for 2:00:00 on cassandra-dev2001.codfw.wmnet with reason: host reimage |
[production] |
| 10:16 |
<mvernon@cumin2003> |
START - Cookbook sre.hosts.reimage for host aqs1019.eqiad.wmnet with OS bookworm |
[production] |
| 10:15 |
<btullis@deploy1003> |
helmfile [staging] DONE helmfile.d/services/eventgate-analytics: sync |
[production] |
| 10:15 |
<btullis@deploy1003> |
helmfile [staging] START helmfile.d/services/eventgate-analytics: sync |
[production] |
| 10:12 |
<marostegui@cumin1003> |
DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 1:00:00 on 24 hosts with reason: Changing sanitarium master in s8 |
[production] |
| 10:11 |
<marostegui> |
Move s8 sanitarium from db1167 to db1281 T434778 |
[production] |
| 10:09 |
<ayounsi@cumin1003> |
END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on netflow2005.codfw.wmnet with reason: host reimage |
[production] |
| 10:04 |
<wm-bot2> |
Deployment completed: https://github.com/cluebotng/component-configs/actions/runs/33742087870 (https://github.com/cluebotng/component-configs/commits/781284a9a6b62d000c7c570fb34d20bbe21933a0) |
[tools.cluebotng] |
| 10:03 |
<ayounsi@cumin1003> |
START - Cookbook sre.hosts.downtime for 2:00:00 on netflow2005.codfw.wmnet with reason: host reimage |
[production] |
| 09:56 |
<elukey@cumin1003> |
START - Cookbook sre.hosts.reimage for host cassandra-dev2001.codfw.wmnet with OS bookworm |
[production] |