1-50 of 10000 results (40ms)
2026-09-03 ยง
21:03 <eevans@cumin1003> END (FAIL) - Cookbook sre.hardware.upgrade-firmware (exit_code=99) upgrade firmware for hosts aqs1021.eqiad.wmnet [production]
20:55 <arlolra@deploy1003> Finished scap sync-world: Backport for [[gerrit:1335109|Add exclusions to Apple app site association file (T435363)]] (duration: 12m 24s) [production]
20:52 <eevans@cumin1003> START - Cookbook sre.hardware.upgrade-firmware upgrade firmware for hosts aqs1021.eqiad.wmnet [production]
20:52 <eevans@cumin1003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host aqs1021.eqiad.wmnet with OS bookworm [production]
20:50 <arlolra@deploy1003> arlolra, tsev: Continuing with deployment [production]
20:46 <arlolra@deploy1003> arlolra, tsev: Backport for [[gerrit:1335109|Add exclusions to Apple app site association file (T435363)]] synced to the testservers (see https://wikitech.wikimedia.org/wiki/Mwdebug). Changes can now be verified there. [production]
20:44 <andrew@cumin1003> END (PASS) - Cookbook sre.hosts.reboot-single (exit_code=0) for host cloudcephosd1047.eqiad.wmnet [production]
20:42 <arlolra@deploy1003> Started scap sync-world: Backport for [[gerrit:1335109|Add exclusions to Apple app site association file (T435363)]] [production]
20:41 <cmooney@cumin1004> END (PASS) - Cookbook sre.network.tls (exit_code=0) for network device lsw1-a3-eqiad [production]
20:40 <cmooney@cumin1004> START - Cookbook sre.network.tls for network device lsw1-a3-eqiad [production]
20:38 <arlolra@deploy1003> Finished scap sync-world: Backport for [[gerrit:1334777|prv: Enable parsoid rendering for more wikisource wikis (T436919)]] (duration: 10m 23s) [production]
20:34 <arlolra@deploy1003> arlolra, jgiannelos: Continuing with deployment [production]
20:33 <andrew@cumin1003> START - Cookbook sre.hosts.reboot-single for host cloudcephosd1047.eqiad.wmnet [production]
20:33 <eevans@cumin1003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on aqs1021.eqiad.wmnet with reason: host reimage [production]
20:32 <arlolra@deploy1003> arlolra, jgiannelos: Backport for [[gerrit:1334777|prv: Enable parsoid rendering for more wikisource wikis (T436919)]] synced to the testservers (see https://wikitech.wikimedia.org/wiki/Mwdebug). Changes can now be verified there. [production]
20:31 <andrew@cumin1003> END (PASS) - Cookbook sre.hosts.reboot-single (exit_code=0) for host cloudcephosd1046.eqiad.wmnet [production]
20:28 <eevans@cumin1003> START - Cookbook sre.hosts.downtime for 2:00:00 on aqs1021.eqiad.wmnet with reason: host reimage [production]
20:28 <arlolra@deploy1003> Started scap sync-world: Backport for [[gerrit:1334777|prv: Enable parsoid rendering for more wikisource wikis (T436919)]] [production]
20:23 <catrope@deploy1003> Finished scap sync-world: Backport for [[gerrit:1334268|Email confirmation A/A test: make registration cutoff consistent (T435135)]], [[gerrit:1335073|Instrumentation for email confirmation upfront enforcement A/A test (T435135)]], [[gerrit:1335074|Email confirmation A/A: check creation wiki, centralize logic (T435135)]] (duration: 13m 41s) [production]
20:20 <andrew@cumin1003> START - Cookbook sre.hosts.reboot-single for host cloudcephosd1046.eqiad.wmnet [production]
20:16 <catrope@deploy1003> catrope: Continuing with deployment [production]
20:14 <eevans@cumin1003> START - Cookbook sre.hosts.reimage for host aqs1021.eqiad.wmnet with OS bookworm [production]
20:13 <catrope@deploy1003> catrope: Backport for [[gerrit:1334268|Email confirmation A/A test: make registration cutoff consistent (T435135)]], [[gerrit:1335073|Instrumentation for email confirmation upfront enforcement A/A test (T435135)]], [[gerrit:1335074|Email confirmation A/A: check creation wiki, centralize logic (T435135)]] synced to the testservers (see https://wikitech.wikimedia.org/wiki/Mwdebug). Changes can now be [production]
20:11 <eevans@cumin1003> END (PASS) - Cookbook sre.hardware.upgrade-firmware (exit_code=0) upgrade firmware for hosts aqs1021.eqiad.wmnet [production]
20:11 <eevans@cumin1003> END (PASS) - Cookbook sre.hosts.reboot-single (exit_code=0) for host aqs1021.eqiad.wmnet [production]
20:09 <catrope@deploy1003> Started scap sync-world: Backport for [[gerrit:1334268|Email confirmation A/A test: make registration cutoff consistent (T435135)]], [[gerrit:1335073|Instrumentation for email confirmation upfront enforcement A/A test (T435135)]], [[gerrit:1335074|Email confirmation A/A: check creation wiki, centralize logic (T435135)]] [production]
20:00 <eevans@cumin1003> START - Cookbook sre.hosts.reboot-single for host aqs1021.eqiad.wmnet [production]
19:49 <eevans@cumin1003> START - Cookbook sre.hardware.upgrade-firmware upgrade firmware for hosts aqs1021.eqiad.wmnet [production]
19:19 <swfrench@deploy1003> Finished scap sync-world: Noop deployment to validate pretrain logstash check configuration - T435419 (duration: 02m 59s) [production]
19:16 <swfrench@deploy1003> Started scap sync-world: Noop deployment to validate pretrain logstash check configuration - T435419 [production]
19:01 <andrew@cumin1003> END (FAIL) - Cookbook sre.hardware.upgrade-firmware (exit_code=99) upgrade firmware for hosts cloudcephosd1045.eqiad.wmnet [production]
19:01 <andrew@cumin1003> START - Cookbook sre.hardware.upgrade-firmware upgrade firmware for hosts cloudcephosd1045.eqiad.wmnet [production]
19:00 <elukey@cumin1003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host cassandra-dev2002.codfw.wmnet with OS bookworm [production]
18:57 <andrew@cumin1003> END (FAIL) - Cookbook sre.hardware.upgrade-firmware (exit_code=99) upgrade firmware for hosts cloudcephosd1045.eqiad.wmnet [production]
18:57 <andrew@cumin1003> START - Cookbook sre.hardware.upgrade-firmware upgrade firmware for hosts cloudcephosd1045.eqiad.wmnet [production]
18:39 <elukey@cumin1003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on cassandra-dev2002.codfw.wmnet with reason: host reimage [production]
18:34 <elukey@cumin1003> START - Cookbook sre.hosts.downtime for 2:00:00 on cassandra-dev2002.codfw.wmnet with reason: host reimage [production]
18:20 <dancy@deploy1003> rebuilt and synchronized wikiversions files: group2 to 1.47.0-wmf.18 refs T430837 [production]
18:16 <elukey@cumin1003> START - Cookbook sre.hosts.reimage for host cassandra-dev2002.codfw.wmnet with OS bookworm [production]
17:55 <andrew@cumin1003> END (FAIL) - Cookbook sre.hardware.upgrade-firmware (exit_code=99) upgrade firmware for hosts cloudcephosd1045.eqiad.wmnet [production]
17:55 <andrew@cumin1003> START - Cookbook sre.hardware.upgrade-firmware upgrade firmware for hosts cloudcephosd1045.eqiad.wmnet [production]
17:54 <andrew@cumin1003> END (PASS) - Cookbook sre.hardware.upgrade-firmware (exit_code=0) upgrade firmware for hosts cloudcephosd1045.eqiad.wmnet [production]
17:54 <andrew@cumin1003> START - Cookbook sre.hardware.upgrade-firmware upgrade firmware for hosts cloudcephosd1045.eqiad.wmnet [production]
17:54 <andrew@cumin1003> END (FAIL) - Cookbook sre.hardware.upgrade-firmware (exit_code=99) upgrade firmware for hosts cloudcephosd1045.eqiad.wmnet [production]
17:53 <andrew@cumin1003> START - Cookbook sre.hardware.upgrade-firmware upgrade firmware for hosts cloudcephosd1045.eqiad.wmnet [production]
17:52 <andrew@cumin1003> END (FAIL) - Cookbook sre.hardware.upgrade-firmware (exit_code=99) upgrade firmware for hosts cloudcephosd1045.eqiad.wmnet [production]
17:46 <jforrester@deploy1003> helmfile [eqiad] DONE helmfile.d/services/wikifunctions: apply [production]
17:46 <jforrester@deploy1003> helmfile [eqiad] START helmfile.d/services/wikifunctions: apply [production]
17:46 <jforrester@deploy1003> helmfile [codfw] DONE helmfile.d/services/wikifunctions: apply [production]
17:46 <swfrench@deploy1003> helmfile [eqiad] DONE helmfile.d/services/rest-gateway: apply [production]