251-300 of 10000 results (45ms)
2026-09-09 ยง
13:53 <btullis@deploy1003> helmfile [ml-staging-codfw] START helmfile.d/admin 'apply'. [production]
13:43 <btullis@deploy1003> helmfile [ml-serve-eqiad] DONE helmfile.d/admin 'apply'. [production]
13:42 <btullis@deploy1003> helmfile [ml-serve-eqiad] START helmfile.d/admin 'apply'. [production]
13:29 <moritzm> pruned obsolete Bullseye image dispatch from the docker registry T416452 [production]
13:28 <sukhe@cumin1003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on sretest2013.codfw.wmnet with reason: host reimage [production]
13:26 <cmooney@cumin1004> END (PASS) - Cookbook sre.network.tls (exit_code=0) for network device lsw1-b7-eqiad [production]
13:25 <cmooney@cumin1004> START - Cookbook sre.network.tls for network device lsw1-b7-eqiad [production]
13:24 <sukhe@cumin1003> START - Cookbook sre.hosts.downtime for 2:00:00 on sretest2013.codfw.wmnet with reason: host reimage [production]
13:22 <ladsgroup@deploy1003> helmfile [staging] DONE helmfile.d/services/thumbor: apply [production]
13:22 <ladsgroup@deploy1003> helmfile [staging] START helmfile.d/services/thumbor: apply [production]
13:17 <cmooney@cumin1004> END (PASS) - Cookbook sre.network.tls (exit_code=0) for network device lsw1-a4-eqiad [production]
13:17 <cmooney@cumin1004> START - Cookbook sre.network.tls for network device lsw1-a4-eqiad [production]
13:15 <sbisson@deploy1003> Finished scap sync-world: Backport for [[gerrit:1334945|ArticleGuidance: Add the redirect configuration keys (T434487)]], [[gerrit:1337983|Replace experiment with instrument and config-driven redirect (T434487)]] (duration: 10m 15s) [production]
13:14 <elukey@cumin1003> END (ERROR) - Cookbook sre.hosts.reimage (exit_code=97) for host cloudcephosd1055.eqiad.wmnet with OS trixie [production]
13:11 <sukhe@cumin1003> START - Cookbook sre.hosts.reimage for host sretest2013.codfw.wmnet with OS trixie [production]
13:10 <sbisson@deploy1003> sbisson: Continuing with deployment [production]
13:09 <sbisson@deploy1003> sbisson: Backport for [[gerrit:1334945|ArticleGuidance: Add the redirect configuration keys (T434487)]], [[gerrit:1337983|Replace experiment with instrument and config-driven redirect (T434487)]] synced to the testservers (see https://wikitech.wikimedia.org/wiki/Mwdebug). Changes can now be verified there. [production]
13:04 <sbisson@deploy1003> Started scap sync-world: Backport for [[gerrit:1334945|ArticleGuidance: Add the redirect configuration keys (T434487)]], [[gerrit:1337983|Replace experiment with instrument and config-driven redirect (T434487)]] [production]
13:02 <jmm@cumin2003> DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 7 days, 0:00:00 on ldap-rw[1001,2001].wikimedia.org with reason: work in progress [production]
12:49 <elukey@cumin1003> START - Cookbook sre.hosts.reimage for host cloudcephosd1055.eqiad.wmnet with OS trixie [production]
12:48 <btullis@deploy1003> helmfile [ml-serve-codfw] DONE helmfile.d/admin 'apply'. [production]
12:46 <btullis@deploy1003> helmfile [ml-serve-codfw] START helmfile.d/admin 'apply'. [production]
12:46 <ladsgroup@deploy1003> Finished scap sync-world: Backport for [[gerrit:1338151|core-Namespaces.php: Disallow indexing on talk namespaces on ukwiki (T437409)]] (duration: 14m 39s) [production]
12:41 <ladsgroup@deploy1003> tryvix1509, ladsgroup: Continuing with deployment [production]
12:35 <ladsgroup@deploy1003> tryvix1509, ladsgroup: Backport for [[gerrit:1338151|core-Namespaces.php: Disallow indexing on talk namespaces on ukwiki (T437409)]] synced to the testservers (see https://wikitech.wikimedia.org/wiki/Mwdebug). Changes can now be verified there. [production]
12:31 <ladsgroup@deploy1003> Started scap sync-world: Backport for [[gerrit:1338151|core-Namespaces.php: Disallow indexing on talk namespaces on ukwiki (T437409)]] [production]
12:16 <ladsgroup@deploy1003> helmfile [staging] DONE helmfile.d/services/thumbor: apply [production]
12:16 <ladsgroup@deploy1003> helmfile [staging] START helmfile.d/services/thumbor: apply [production]
11:53 <urbanecm@deploy1003> Finished scap sync-world: Backport for [[gerrit:1338031|[Growth] Enable iterative Add Link task pool population on all wikis (T392944)]] (duration: 21m 58s) [production]
11:48 <urbanecm@deploy1003> urbanecm: Continuing with deployment [production]
11:35 <urbanecm@deploy1003> urbanecm: Backport for [[gerrit:1338031|[Growth] Enable iterative Add Link task pool population on all wikis (T392944)]] synced to the testservers (see https://wikitech.wikimedia.org/wiki/Mwdebug). Changes can now be verified there. [production]
11:31 <urbanecm@deploy1003> Started scap sync-world: Backport for [[gerrit:1338031|[Growth] Enable iterative Add Link task pool population on all wikis (T392944)]] [production]
10:55 <marostegui@cumin1003> END (PASS) - Cookbook sre.mysql.pool (exit_code=0) pool db2196: Repooling db2196 [production]
10:47 <moritzm> pruned obsolete Bullseye images nodejs12-slim/nodejs12-devel/nodejs14-slim/nodejs16-slim from the docker registry T416452 [production]
10:43 <moritzm> installing Bird security updates [production]
10:38 <marostegui@cumin1003> END (PASS) - Cookbook sre.mysql.pool (exit_code=0) pool db1260: Repooling after cloning [production]
10:09 <marostegui@cumin1003> START - Cookbook sre.mysql.pool pool db2196: Repooling db2196 [production]
10:07 <sfaci@deploy1003> helmfile [dse-k8s-eqiad] DONE helmfile.d/dse-k8s-services/test-kitchen-next: apply [production]
10:06 <sfaci@deploy1003> helmfile [dse-k8s-eqiad] START helmfile.d/dse-k8s-services/test-kitchen-next: apply [production]
09:55 <moritzm> pruned obsolete Bullseye images openjdk-8-jdk/openjdk-8-jre/openjdk-11-jre/openjdk-11-jdk from the docker registry T416452 [production]
09:53 <marostegui@cumin1003> START - Cookbook sre.mysql.pool pool db1260: Repooling after cloning [production]
09:52 <cmooney@cumin1004> END (PASS) - Cookbook sre.network.tls (exit_code=0) for network device lsw1-d3-codfw [production]
09:52 <cmooney@cumin1004> START - Cookbook sre.network.tls for network device lsw1-d3-codfw [production]
09:51 <brouberol@deploy1003> helmfile [dse-k8s-eqiad] DONE helmfile.d/services/mediawiki-dumps-legacy: apply [production]
09:43 <brouberol@deploy1003> helmfile [dse-k8s-eqiad] START helmfile.d/services/mediawiki-dumps-legacy: apply [production]
09:28 <btullis@cumin1003> END (FAIL) - Cookbook sre.hardware.upgrade-firmware (exit_code=99) upgrade firmware for hosts an-worker1198.eqiad.wmnet [production]
09:27 <btullis@cumin1003> START - Cookbook sre.hardware.upgrade-firmware upgrade firmware for hosts an-worker1198.eqiad.wmnet [production]
09:03 <brouberol@deploy1003> helmfile [dse-k8s-codfw] DONE helmfile.d/admin 'apply'. [production]
09:02 <brouberol@deploy1003> helmfile [dse-k8s-codfw] START helmfile.d/admin 'apply'. [production]
09:01 <cmooney@cumin1004> DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on 16 hosts with reason: upgrade ssw1-a1-eqiad [production]