151-200 of 10000 results (48ms)
2026-09-09 §
11:48 <urbanecm@deploy1003> urbanecm: Continuing with deployment [production]
11:35 <urbanecm@deploy1003> urbanecm: Backport for [[gerrit:1338031|[Growth] Enable iterative Add Link task pool population on all wikis (T392944)]] synced to the testservers (see https://wikitech.wikimedia.org/wiki/Mwdebug). Changes can now be verified there. [production]
11:31 <urbanecm@deploy1003> Started scap sync-world: Backport for [[gerrit:1338031|[Growth] Enable iterative Add Link task pool population on all wikis (T392944)]] [production]
10:55 <marostegui@cumin1003> END (PASS) - Cookbook sre.mysql.pool (exit_code=0) pool db2196: Repooling db2196 [production]
10:47 <moritzm> pruned obsolete Bullseye images nodejs12-slim/nodejs12-devel/nodejs14-slim/nodejs16-slim from the docker registry T416452 [production]
10:43 <moritzm> installing Bird security updates [production]
10:38 <marostegui@cumin1003> END (PASS) - Cookbook sre.mysql.pool (exit_code=0) pool db1260: Repooling after cloning [production]
10:09 <marostegui@cumin1003> START - Cookbook sre.mysql.pool pool db2196: Repooling db2196 [production]
10:07 <sfaci@deploy1003> helmfile [dse-k8s-eqiad] DONE helmfile.d/dse-k8s-services/test-kitchen-next: apply [production]
10:06 <sfaci@deploy1003> helmfile [dse-k8s-eqiad] START helmfile.d/dse-k8s-services/test-kitchen-next: apply [production]
09:55 <moritzm> pruned obsolete Bullseye images openjdk-8-jdk/openjdk-8-jre/openjdk-11-jre/openjdk-11-jdk from the docker registry T416452 [production]
09:53 <marostegui@cumin1003> START - Cookbook sre.mysql.pool pool db1260: Repooling after cloning [production]
09:52 <cmooney@cumin1004> END (PASS) - Cookbook sre.network.tls (exit_code=0) for network device lsw1-d3-codfw [production]
09:52 <cmooney@cumin1004> START - Cookbook sre.network.tls for network device lsw1-d3-codfw [production]
09:51 <brouberol@deploy1003> helmfile [dse-k8s-eqiad] DONE helmfile.d/services/mediawiki-dumps-legacy: apply [production]
09:43 <brouberol@deploy1003> helmfile [dse-k8s-eqiad] START helmfile.d/services/mediawiki-dumps-legacy: apply [production]
09:28 <btullis@cumin1003> END (FAIL) - Cookbook sre.hardware.upgrade-firmware (exit_code=99) upgrade firmware for hosts an-worker1198.eqiad.wmnet [production]
09:27 <btullis@cumin1003> START - Cookbook sre.hardware.upgrade-firmware upgrade firmware for hosts an-worker1198.eqiad.wmnet [production]
09:03 <brouberol@deploy1003> helmfile [dse-k8s-codfw] DONE helmfile.d/admin 'apply'. [production]
09:02 <brouberol@deploy1003> helmfile [dse-k8s-codfw] START helmfile.d/admin 'apply'. [production]
09:01 <cmooney@cumin1004> DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on 16 hosts with reason: upgrade ssw1-a1-eqiad [production]
08:58 <cmooney@cumin1004> DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on 22 hosts with reason: upgrade ssw1-a1-eqiad [production]
08:56 <brouberol@deploy1003> helmfile [dse-k8s-eqiad] DONE helmfile.d/admin 'apply'. [production]
08:55 <brouberol@deploy1003> helmfile [dse-k8s-eqiad] START helmfile.d/admin 'apply'. [production]
08:49 <cmooney@cumin1004> END (PASS) - Cookbook sre.network.tls (exit_code=0) for network device lsw1-d3-codfw [production]
08:49 <cmooney@cumin1004> START - Cookbook sre.network.tls for network device lsw1-d3-codfw [production]
08:48 <cmooney@cumin1004> END (PASS) - Cookbook sre.network.tls (exit_code=0) for network device ssw1-d1-codfw [production]
08:48 <cmooney@cumin1004> START - Cookbook sre.network.tls for network device ssw1-d1-codfw [production]
08:36 <marostegui@cumin1003> DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 12:00:00 on db1155.eqiad.wmnet with reason: Cloning sanitarium [production]
08:30 <brouberol@dns1004> END - running authdns-update [production]
08:28 <moritzm> pruned obsolete Bullseye image golang1.15 from the docker registry T416452 [production]
08:28 <brouberol@dns1004> START - running authdns-update [production]
08:06 <marostegui@cumin1003> END (PASS) - Cookbook sre.mysql.depool (exit_code=0) depool db1260: Needs to clone another host from this one [production]
08:06 <marostegui@cumin1003> START - Cookbook sre.mysql.depool depool db1260: Needs to clone another host from this one [production]
08:03 <marostegui@cumin1003> DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 12:00:00 on db1260.eqiad.wmnet with reason: Cloning sanitarium [production]
08:01 <btullis@cumin1003> END (FAIL) - Cookbook sre.hardware.upgrade-firmware (exit_code=99) upgrade firmware for hosts an-worker1198.eqiad.wmnet [production]
08:01 <btullis@cumin1003> START - Cookbook sre.hardware.upgrade-firmware upgrade firmware for hosts an-worker1198.eqiad.wmnet [production]
08:01 <btullis@cumin1003> END (FAIL) - Cookbook sre.hardware.upgrade-firmware (exit_code=99) upgrade firmware for hosts an-worker1198.eqiad.wmnet [production]
08:00 <btullis@cumin1003> START - Cookbook sre.hardware.upgrade-firmware upgrade firmware for hosts an-worker1198.eqiad.wmnet [production]
07:40 <chlod> UTC morning backport window done [production]
07:37 <chlod@deploy1003> Finished scap sync-world: Backport for [[gerrit:1335706|thwikibooks: update tagline and wordmark (T436426)]] (duration: 21m 36s) [production]
07:32 <chlod@deploy1003> chlod, hamishz: Continuing with deployment [production]
07:20 <chlod@deploy1003> chlod, hamishz: Backport for [[gerrit:1335706|thwikibooks: update tagline and wordmark (T436426)]] synced to the testservers (see https://wikitech.wikimedia.org/wiki/Mwdebug). Changes can now be verified there. [production]
07:15 <chlod@deploy1003> Started scap sync-world: Backport for [[gerrit:1335706|thwikibooks: update tagline and wordmark (T436426)]] [production]
02:08 <mwpresync@deploy1003> Finished scap build-images: Publishing wmf/next image (duration: 07m 45s) [production]
02:00 <mwpresync@deploy1003> Started scap build-images: Publishing wmf/next image [production]
00:09 <eevans@cumin1003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host aqs1025.eqiad.wmnet with OS bookworm [production]
2026-09-08 §
23:51 <eevans@cumin1003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on aqs1025.eqiad.wmnet with reason: host reimage [production]
23:48 <eevans@cumin1003> START - Cookbook sre.hosts.downtime for 2:00:00 on aqs1025.eqiad.wmnet with reason: host reimage [production]
23:39 <eevans@cumin1003> START - Cookbook sre.hosts.reimage for host aqs1025.eqiad.wmnet with OS bookworm [production]