1-50 of 10000 results (3ms)
2026-09-04 ยง
10:49 <jmm@cumin1004> END (PASS) - Cookbook sre.dns.wipe-cache (exit_code=0) pki2003.codfw.wmnet on all recursors [production]
10:49 <jmm@cumin1004> START - Cookbook sre.dns.wipe-cache pki2003.codfw.wmnet on all recursors [production]
10:49 <jmm@cumin1004> END (PASS) - Cookbook sre.dns.netbox (exit_code=0) [production]
10:49 <jmm@cumin1004> END (PASS) - Cookbook sre.puppet.sync-netbox-hiera (exit_code=0) generate netbox hiera data: "Triggered by cookbooks.sre.dns.netbox: Add records for VM pki2003.codfw.wmnet - jmm@cumin1004" [production]
10:49 <jmm@cumin1004> START - Cookbook sre.puppet.sync-netbox-hiera generate netbox hiera data: "Triggered by cookbooks.sre.dns.netbox: Add records for VM pki2003.codfw.wmnet - jmm@cumin1004" [production]
10:44 <jmm@cumin1004> START - Cookbook sre.dns.netbox [production]
10:44 <jmm@cumin1004> START - Cookbook sre.ganeti.makevm for new host pki2003.codfw.wmnet [production]
10:29 <btullis@deploy1003> Finished scap sync-world: Incorporating changes from https://gerrit.wikimedia.org/r/c/operations/dumps/+/1334846 to mediawiki-cli (duration: 41m 14s) [production]
10:00 <marostegui@cumin1003> END (PASS) - Cookbook sre.mysql.decommission (exit_code=0) [production]
10:00 <marostegui@cumin1003> Removing db1182 from zarcillo T434869 [production]
10:00 <marostegui@cumin1003> END (FAIL) - Cookbook sre.hosts.decommission (exit_code=1) for hosts db1182.eqiad.wmnet [production]
10:00 <marostegui@cumin1003> END (PASS) - Cookbook sre.dns.netbox (exit_code=0) [production]
09:57 <marostegui@cumin1003> START - Cookbook sre.dns.netbox [production]
09:57 <btullis@deploy1003> Started scap sync-world: Incorporating changes from https://gerrit.wikimedia.org/r/c/operations/dumps/+/1334846 to mediawiki-cli [production]
09:53 <marostegui@cumin1003> START - Cookbook sre.hosts.decommission for hosts db1182.eqiad.wmnet [production]
09:53 <marostegui@cumin1003> START - Cookbook sre.mysql.decommission [production]
09:53 <marostegui@cumin1003> END (ERROR) - Cookbook sre.mysql.decommission (exit_code=1) [production]
09:53 <marostegui@cumin1003> START - Cookbook sre.mysql.decommission [production]
09:51 <marostegui@cumin1003> END (PASS) - Cookbook sre.hosts.decommission (exit_code=0) for hosts db1182.eqiad.wmnet [production]
09:51 <marostegui@cumin1003> END (PASS) - Cookbook sre.dns.netbox (exit_code=0) [production]
09:51 <marostegui@cumin1003> END (PASS) - Cookbook sre.puppet.sync-netbox-hiera (exit_code=0) generate netbox hiera data: "Triggered by cookbooks.sre.dns.netbox: db1182.eqiad.wmnet decommissioned, removing all IPs except the asset tag one - marostegui@cumin1003" [production]
09:50 <marostegui@cumin1003> START - Cookbook sre.puppet.sync-netbox-hiera generate netbox hiera data: "Triggered by cookbooks.sre.dns.netbox: db1182.eqiad.wmnet decommissioned, removing all IPs except the asset tag one - marostegui@cumin1003" [production]
09:46 <marostegui@cumin1003> START - Cookbook sre.dns.netbox [production]
09:45 <marostegui@cumin1003> dbctl commit (dc=all): 'Remove db1182 from dbctl T434869', diff saved to https://phabricator.wikimedia.org/P96346 and previous config saved to /var/cache/conftool/dbconfig/20260904-094527-marostegui.json [production]
09:41 <marostegui@cumin1003> START - Cookbook sre.hosts.decommission for hosts db1182.eqiad.wmnet [production]
09:39 <jmm@cumin1004> END (PASS) - Cookbook sre.ganeti.makevm (exit_code=0) for new host pki1003.eqiad.wmnet [production]
09:39 <jmm@cumin1004> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host pki1003.eqiad.wmnet with OS trixie [production]
09:32 <marostegui@cumin1003> END (PASS) - Cookbook sre.mysql.depool (exit_code=0) depool db1182: Decommissioning [production]
09:31 <marostegui@cumin1003> START - Cookbook sre.mysql.depool depool db1182: Decommissioning [production]
09:23 <jmm@cumin1004> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on pki1003.eqiad.wmnet with reason: host reimage [production]
09:18 <elukey@cumin1003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host cassandra-dev2003.codfw.wmnet with OS bookworm [production]
09:17 <jmm@cumin1004> START - Cookbook sre.hosts.downtime for 2:00:00 on pki1003.eqiad.wmnet with reason: host reimage [production]
09:09 <ayounsi@cumin1003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host netflow5003.eqsin.wmnet with OS trixie [production]
09:02 <jmm@cumin1004> START - Cookbook sre.hosts.reimage for host pki1003.eqiad.wmnet with OS trixie [production]
09:00 <jmm@cumin1004> END (PASS) - Cookbook sre.puppet.sync-netbox-hiera (exit_code=0) generate netbox hiera data: "Triggered by cookbooks.sre.ganeti.makevm: created new VM pki1003.eqiad.wmnet - jmm@cumin1004" [production]
09:00 <jmm@cumin1004> START - Cookbook sre.puppet.sync-netbox-hiera generate netbox hiera data: "Triggered by cookbooks.sre.ganeti.makevm: created new VM pki1003.eqiad.wmnet - jmm@cumin1004" [production]
08:59 <jmm@cumin1004> END (PASS) - Cookbook sre.dns.wipe-cache (exit_code=0) pki1003.eqiad.wmnet on all recursors [production]
08:59 <jmm@cumin1004> START - Cookbook sre.dns.wipe-cache pki1003.eqiad.wmnet on all recursors [production]
08:59 <jmm@cumin1004> END (PASS) - Cookbook sre.dns.netbox (exit_code=0) [production]
08:59 <jmm@cumin1004> END (PASS) - Cookbook sre.puppet.sync-netbox-hiera (exit_code=0) generate netbox hiera data: "Triggered by cookbooks.sre.dns.netbox: Add records for VM pki1003.eqiad.wmnet - jmm@cumin1004" [production]
08:59 <jmm@cumin1004> START - Cookbook sre.puppet.sync-netbox-hiera generate netbox hiera data: "Triggered by cookbooks.sre.dns.netbox: Add records for VM pki1003.eqiad.wmnet - jmm@cumin1004" [production]
08:58 <elukey@cumin1003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on cassandra-dev2003.codfw.wmnet with reason: host reimage [production]
08:55 <jmm@cumin1004> START - Cookbook sre.dns.netbox [production]
08:55 <jmm@cumin1004> START - Cookbook sre.ganeti.makevm for new host pki1003.eqiad.wmnet [production]
08:54 <elukey@cumin1003> START - Cookbook sre.hosts.downtime for 2:00:00 on cassandra-dev2003.codfw.wmnet with reason: host reimage [production]
08:48 <ayounsi@cumin1003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on netflow5003.eqsin.wmnet with reason: host reimage [production]
08:45 <ayounsi@cumin1003> START - Cookbook sre.hosts.downtime for 2:00:00 on netflow5003.eqsin.wmnet with reason: host reimage [production]
08:40 <btullis@deploy1003> Finished scap sync-world: Trying again for T436913 (duration: 34m 26s) [production]
08:35 <elukey@cumin1003> START - Cookbook sre.hosts.reimage for host cassandra-dev2003.codfw.wmnet with OS bookworm [production]
08:35 <elukey@cumin1003> END (PASS) - Cookbook sre.hardware.upgrade-firmware (exit_code=0) upgrade firmware for hosts cassandra-dev2003.codfw.wmnet [production]