201-250 of 10000 results (20ms)
2026-08-12 ยง
14:05 <cgoubert@cumin2003> START - Cookbook sre.puppet.sync-netbox-hiera generate netbox hiera data: "Triggered by cookbooks.sre.ganeti.makevm: created new VM rdb-lock1002.eqiad.wmnet - cgoubert@cumin2003" [production]
14:05 <cgoubert@cumin2003> END (PASS) - Cookbook sre.dns.wipe-cache (exit_code=0) rdb-lock1002.eqiad.wmnet on all recursors [production]
14:05 <cgoubert@cumin2003> START - Cookbook sre.dns.wipe-cache rdb-lock1002.eqiad.wmnet on all recursors [production]
14:05 <cgoubert@cumin2003> END (PASS) - Cookbook sre.dns.netbox (exit_code=0) [production]
14:05 <cgoubert@cumin2003> END (PASS) - Cookbook sre.puppet.sync-netbox-hiera (exit_code=0) generate netbox hiera data: "Triggered by cookbooks.sre.dns.netbox: Add records for VM rdb-lock1002.eqiad.wmnet - cgoubert@cumin2003" [production]
14:04 <jforrester@deploy1003> helmfile [staging] DONE helmfile.d/services/wikifunctions: apply [production]
14:04 <kharlan@deploy1003> kharlan: Continuing with deployment [production]
14:04 <jforrester@deploy1003> helmfile [staging] START helmfile.d/services/wikifunctions: apply [production]
14:03 <cgoubert@cumin2003> END (PASS) - Cookbook sre.ganeti.makevm (exit_code=0) for new host rdb-lock2001.codfw.wmnet [production]
14:03 <cgoubert@cumin2003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host rdb-lock2001.codfw.wmnet with OS trixie [production]
14:02 <cgoubert@cumin2003> START - Cookbook sre.puppet.sync-netbox-hiera generate netbox hiera data: "Triggered by cookbooks.sre.dns.netbox: Add records for VM rdb-lock1002.eqiad.wmnet - cgoubert@cumin2003" [production]
14:02 <bking@cumin2003> START - Cookbook sre.hosts.downtime for 2:00:00 on apifeatureusage2001.codfw.wmnet with reason: host reimage [production]
14:00 <jmm@cumin2003> START - Cookbook sre.ganeti.drain-node for draining ganeti node ganeti2030.codfw.wmnet [production]
13:59 <jmm@cumin2003> END (PASS) - Cookbook sre.ganeti.drain-node (exit_code=0) for draining ganeti node ganeti2029.codfw.wmnet [production]
13:58 <jmm@cumin2003> END (PASS) - Cookbook sre.hosts.reboot-single (exit_code=0) for host ganeti2029.codfw.wmnet [production]
13:58 <cgoubert@cumin2003> START - Cookbook sre.dns.netbox [production]
13:58 <cgoubert@cumin2003> START - Cookbook sre.ganeti.makevm for new host rdb-lock1002.eqiad.wmnet [production]
13:56 <btullis@cumin1003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on an-presto1019.eqiad.wmnet with reason: host reimage [production]
13:54 <btullis@cumin1003> DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 4:00:00 on archiva1002.wikimedia.org with reason: Upgrading in-place [production]
13:53 <cgoubert@cumin2003> END (PASS) - Cookbook sre.ganeti.makevm (exit_code=0) for new host rdb-lock1001.eqiad.wmnet [production]
13:53 <cgoubert@cumin2003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host rdb-lock1001.eqiad.wmnet with OS trixie [production]
13:53 <kharlan@deploy1003> kharlan: Backport for [[gerrit:1324707|Backport all changes from wmf/1.47.0-wmf.15]] synced to the testservers (see https://wikitech.wikimedia.org/wiki/Mwdebug). Changes can now be verified there. [production]
13:52 <jmm@cumin2003> START - Cookbook sre.hosts.reboot-single for host ganeti2029.codfw.wmnet [production]
13:52 <btullis@cumin1003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on an-presto1020.eqiad.wmnet with reason: host reimage [production]
13:49 <btullis@cumin1003> START - Cookbook sre.hosts.downtime for 2:00:00 on an-presto1019.eqiad.wmnet with reason: host reimage [production]
13:49 <btullis@cumin1003> START - Cookbook sre.hosts.downtime for 2:00:00 on an-presto1020.eqiad.wmnet with reason: host reimage [production]
13:49 <cgoubert@cumin2003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on rdb-lock2001.codfw.wmnet with reason: host reimage [production]
13:48 <jmm@cumin2003> START - Cookbook sre.ganeti.drain-node for draining ganeti node ganeti2029.codfw.wmnet [production]
13:47 <cwilliams@cumin1003> dbctl commit (dc=all): 'Configuring db1272 for s3 pooling', diff saved to https://phabricator.wikimedia.org/P96021 and previous config saved to /var/cache/conftool/dbconfig/20260812-134732-cwilliams.json [production]
13:44 <bking@cumin2003> START - Cookbook sre.hosts.reimage for host apifeatureusage2001.codfw.wmnet with OS bookworm [production]
13:43 <cgoubert@cumin2003> START - Cookbook sre.hosts.downtime for 2:00:00 on rdb-lock2001.codfw.wmnet with reason: host reimage [production]
13:41 <jmm@cumin2003> END (PASS) - Cookbook sre.ganeti.drain-node (exit_code=0) for draining ganeti node ganeti2028.codfw.wmnet [production]
13:41 <jmm@cumin2003> END (PASS) - Cookbook sre.hosts.reboot-single (exit_code=0) for host ganeti2028.codfw.wmnet [production]
13:41 <btullis@cumin1003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on an-presto1018.eqiad.wmnet with reason: host reimage [production]
13:40 <taavi@cloudcumin1001> END (PASS) - Cookbook wmcs.vps.remove_instance (exit_code=0) for instance cloudinfra-idp-1 [cloudinfra]
13:39 <taavi@cloudcumin1001> START - Cookbook wmcs.vps.remove_instance for instance cloudinfra-idp-1 [cloudinfra]
13:38 <cgoubert@cumin2003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on rdb-lock1001.eqiad.wmnet with reason: host reimage [production]
13:36 <btullis@cumin1003> START - Cookbook sre.hosts.downtime for 2:00:00 on an-presto1018.eqiad.wmnet with reason: host reimage [production]
13:35 <kharlan@deploy1003> Started scap sync-world: Backport for [[gerrit:1324707|Backport all changes from wmf/1.47.0-wmf.15]] [production]
13:35 <jmm@cumin2003> START - Cookbook sre.hosts.reboot-single for host ganeti2028.codfw.wmnet [production]
13:32 <kharlan@deploy1003> Finished scap sync-world: Backport for [[gerrit:1275938|Enable campaignEvents on bdwikimedia (T424016)]] (duration: 07m 35s) [production]
13:32 <btullis@cumin1003> START - Cookbook sre.hosts.reimage for host an-presto1020.eqiad.wmnet with OS bookworm [production]
13:32 <btullis@cumin1003> START - Cookbook sre.hosts.reimage for host an-presto1019.eqiad.wmnet with OS bookworm [production]
13:32 <cgoubert@cumin2003> START - Cookbook sre.hosts.downtime for 2:00:00 on rdb-lock1001.eqiad.wmnet with reason: host reimage [production]
13:28 <dcaro@cloudcumin1001> END (PASS) - Cookbook wmcs.toolforge.component.deploy (exit_code=0) for component misctools-cli [tools]
13:28 <kharlan@deploy1003> kharlan, yahya: Continuing with deployment [production]
13:27 <kharlan@deploy1003> kharlan, yahya: Backport for [[gerrit:1275938|Enable campaignEvents on bdwikimedia (T424016)]] synced to the testservers (see https://wikitech.wikimedia.org/wiki/Mwdebug). Changes can now be verified there. [production]
13:27 <jmm@cumin2003> START - Cookbook sre.ganeti.drain-node for draining ganeti node ganeti2028.codfw.wmnet [production]
13:26 <atsuko@deploy1003> helmfile [dse-k8s-eqiad] DONE helmfile.d/dse-k8s-services/airflow-test-k8s: apply [production]
13:26 <atsuko@deploy1003> helmfile [dse-k8s-eqiad] START helmfile.d/dse-k8s-services/airflow-test-k8s: apply [production]