1-50 of 10000 results (13ms)
2026-08-13 ยง
14:03 <tappof@cumin1003> START - Cookbook sre.hosts.reboot-single for host prometheus2007.codfw.wmnet [production]
14:03 <tappof@cumin1003> END (PASS) - Cookbook sre.hosts.reboot-single (exit_code=0) for host prometheus2005.codfw.wmnet [production]
14:03 <tappof@cumin1003> END (PASS) - Cookbook sre.hosts.reboot-single (exit_code=0) for host prometheus1007.eqiad.wmnet [production]
14:03 <jmm@cumin2003> END (PASS) - Cookbook sre.hosts.reboot-single (exit_code=0) for host netflow1002.eqiad.wmnet [production]
14:02 <bking@cumin2003> START - Cookbook sre.hosts.downtime for 2:00:00 on apifeatureusage1001.eqiad.wmnet with reason: host reimage [production]
14:02 <cgoubert@cumin2003> conftool action : set/pooled=true; selector: dnsdisc=helm-charts.*,name=eqiad [production]
14:01 <cgoubert@cumin2003> END (PASS) - Cookbook sre.hosts.reboot-single (exit_code=0) for host chartmuseum1001.eqiad.wmnet [production]
14:00 <moritzm> installing libxml2 security updates [production]
13:59 <btullis@cumin1003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host an-worker1214.eqiad.wmnet with OS bookworm [production]
13:59 <cmooney@cumin1003> conftool action : set/pooled=yes; selector: name=dns5003.* [production]
13:59 <jmm@cumin2003> START - Cookbook sre.ganeti.drain-node for draining ganeti node ganeti1041.eqiad.wmnet [production]
13:59 <jmm@cumin2003> START - Cookbook sre.hosts.reboot-single for host netflow1002.eqiad.wmnet [production]
13:58 <cmooney@dns3003> END - running authdns-update [production]
13:57 <cgoubert@cumin2003> START - Cookbook sre.hosts.reboot-single for host chartmuseum1001.eqiad.wmnet [production]
13:57 <cgoubert@cumin2003> conftool action : set/pooled=false; selector: dnsdisc=helm-charts.*,name=eqiad [production]
13:57 <bking@cumin2003> END (PASS) - Cookbook sre.elasticsearch.rolling-operation (exit_code=0) Operation.RESTART (1 nodes at a time) for ElasticSearch cluster cloudelastic: cloudelastic cluster restart - bking@cumin2003 [production]
13:57 <cgoubert@cumin2003> conftool action : set/pooled=true; selector: dnsdisc=helm-charts.*,name=codfw [production]
13:56 <cmooney@dns3003> START - running authdns-update [production]
13:56 <jayme@deploy1003> helmfile [codfw] START helmfile.d/admin 'sync'. [production]
13:56 <cmooney@cumin1003> conftool action : set/pooled=yes; selector: name=dns5003.*,service=authdns-update [production]
13:55 <cgoubert@cumin2003> END (PASS) - Cookbook sre.hosts.reboot-single (exit_code=0) for host chartmuseum2001.codfw.wmnet [production]
13:55 <tappof@cumin1003> START - Cookbook sre.hosts.reboot-single for host prometheus1007.eqiad.wmnet [production]
13:55 <cmooney@cumin1003> END (PASS) - Cookbook sre.hosts.reboot-single (exit_code=0) for host dns5003.wikimedia.org [production]
13:55 <tappof@cumin1003> END (PASS) - Cookbook sre.hosts.reboot-single (exit_code=0) for host prometheus1005.eqiad.wmnet [production]
13:51 <cgoubert@cumin2003> START - Cookbook sre.hosts.reboot-single for host chartmuseum2001.codfw.wmnet [production]
13:51 <cgoubert@cumin2003> conftool action : set/pooled=false; selector: dnsdisc=helm-charts.*,name=codfw [production]
13:51 <tappof@cumin1003> START - Cookbook sre.hosts.reboot-single for host prometheus2005.codfw.wmnet [production]
13:51 <bking@cumin2003> START - Cookbook sre.hosts.reimage for host apifeatureusage1001.eqiad.wmnet with OS bookworm [production]
13:50 <tappof@cumin1003> END (PASS) - Cookbook sre.hosts.reboot-single (exit_code=0) for host prometheus7002.magru.wmnet [production]
13:50 <sukhe@cumin1003> END (PASS) - Cookbook sre.hosts.decommission (exit_code=0) for hosts lvs1015.eqiad.wmnet [production]
13:50 <sukhe@cumin1003> END (PASS) - Cookbook sre.dns.netbox (exit_code=0) [production]
13:50 <sukhe@cumin1003> END (PASS) - Cookbook sre.puppet.sync-netbox-hiera (exit_code=0) generate netbox hiera data: "Triggered by cookbooks.sre.dns.netbox: lvs1015.eqiad.wmnet decommissioned, removing all IPs except the asset tag one - sukhe@cumin1003" [production]
13:49 <sukhe@cumin1003> START - Cookbook sre.puppet.sync-netbox-hiera generate netbox hiera data: "Triggered by cookbooks.sre.dns.netbox: lvs1015.eqiad.wmnet decommissioned, removing all IPs except the asset tag one - sukhe@cumin1003" [production]
13:49 <stran@deploy1003> Finished scap sync-world: Backport for [[gerrit:1325480|Guard against malformed headers in SuggestedInvestigationsMatchSignalsAgainstUserJob (T434199)]] (duration: 06m 39s) [production]
13:47 <btullis@cumin1003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on an-worker1215.eqiad.wmnet with reason: host reimage [production]
13:46 <cmooney@cumin1003> START - Cookbook sre.hosts.reboot-single for host dns5003.wikimedia.org [production]
13:46 <jmm@cumin2003> END (PASS) - Cookbook sre.hosts.reboot-single (exit_code=0) for host netflow1003.eqiad.wmnet [production]
13:45 <cmooney@cumin1003> conftool action : set/pooled=no; selector: name=dns5003.* [production]
13:45 <jmm@cumin2003> END (PASS) - Cookbook sre.ganeti.drain-node (exit_code=0) for draining ganeti node ganeti1040.eqiad.wmnet [production]
13:45 <jmm@cumin2003> END (PASS) - Cookbook sre.hosts.reboot-single (exit_code=0) for host ganeti1040.eqiad.wmnet [production]
13:45 <sukhe@cumin1003> START - Cookbook sre.dns.netbox [production]
13:45 <tappof@cumin1003> START - Cookbook sre.hosts.reboot-single for host prometheus1005.eqiad.wmnet [production]
13:45 <stran@deploy1003> stran: Continuing with deployment [production]
13:44 <tappof@cumin1003> START - Cookbook sre.hosts.reboot-single for host prometheus7002.magru.wmnet [production]
13:44 <stran@deploy1003> stran: Backport for [[gerrit:1325480|Guard against malformed headers in SuggestedInvestigationsMatchSignalsAgainstUserJob (T434199)]] synced to the testservers (see https://wikitech.wikimedia.org/wiki/Mwdebug). Changes can now be verified there. [production]
13:44 <tappof@cumin1003> END (PASS) - Cookbook sre.hosts.reboot-single (exit_code=0) for host prometheus6002.drmrs.wmnet [production]
13:43 <Dreamy_Jazz> Running `mwscript-k8s WikimediaAntiAbuse:BackfillAbuseReview.php --wiki=enwiki --start-timestamp="20260805000000" --end-timestamp="20260806000000" --sleep="5" --batch-size="10"` for T434688 [production]
13:43 <btullis@cumin1003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on an-worker1149.eqiad.wmnet with reason: host reimage [production]
13:42 <stran@deploy1003> Started scap sync-world: Backport for [[gerrit:1325480|Guard against malformed headers in SuggestedInvestigationsMatchSignalsAgainstUserJob (T434199)]] [production]
13:41 <jmm@cumin2003> START - Cookbook sre.hosts.reboot-single for host netflow1003.eqiad.wmnet [production]