201-250 of 10000 results (120ms)
2026-08-13 ยง
14:16 <cgoubert@cumin2003> START - Cookbook sre.k8s.pool-depool-node depool for host wikikube-ctrl1003.eqiad.wmnet [production]
14:16 <cgoubert@cumin2003> END (PASS) - Cookbook sre.k8s.pool-depool-node (exit_code=0) pool for host wikikube-ctrl1002.eqiad.wmnet [production]
14:16 <cgoubert@cumin2003> START - Cookbook sre.k8s.pool-depool-node pool for host wikikube-ctrl1002.eqiad.wmnet [production]
14:15 <tappof@cumin1003> START - Cookbook sre.hosts.reboot-single for host prometheus1008.eqiad.wmnet [production]
14:14 <tappof@cumin1003> END (PASS) - Cookbook sre.hosts.reboot-single (exit_code=0) for host prometheus1006.eqiad.wmnet [production]
14:12 <tappof@cumin1003> START - Cookbook sre.hosts.reboot-single for host prometheus2006.codfw.wmnet [production]
14:12 <jmm@cumin2003> END (PASS) - Cookbook sre.hosts.reboot-single (exit_code=0) for host debmonitor2003.codfw.wmnet [production]
14:11 <tappof@cumin1003> END (PASS) - Cookbook sre.hosts.reboot-single (exit_code=0) for host prometheus2007.codfw.wmnet [production]
14:11 <jmm@cumin2003> END (PASS) - Cookbook sre.ganeti.drain-node (exit_code=0) for draining ganeti node ganeti1041.eqiad.wmnet [production]
14:11 <jmm@cumin2003> END (PASS) - Cookbook sre.hosts.reboot-single (exit_code=0) for host ganeti1041.eqiad.wmnet [production]
14:09 <cgoubert@cumin2003> END (PASS) - Cookbook sre.k8s.pool-depool-node (exit_code=0) depool for host wikikube-ctrl1002.eqiad.wmnet [production]
14:09 <cgoubert@cumin2003> START - Cookbook sre.k8s.pool-depool-node depool for host wikikube-ctrl1002.eqiad.wmnet [production]
14:09 <cgoubert@cumin2003> START - Cookbook sre.k8s.reboot-nodes rolling reboot on A:wikikube-master-eqiad [production]
14:09 <bking@cumin2003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on apifeatureusage1001.eqiad.wmnet with reason: host reimage [production]
14:08 <btullis@cumin1003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host an-worker1215.eqiad.wmnet with OS bookworm [production]
14:08 <jmm@cumin2003> START - Cookbook sre.hosts.reboot-single for host debmonitor2003.codfw.wmnet [production]
14:08 <jayme> updated calico to v3.30.7 on wikikube codfw T427400 [production]
14:07 <btullis@cumin1003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host an-worker1149.eqiad.wmnet with OS bookworm [production]
14:06 <jayme@deploy1003> helmfile [codfw] DONE helmfile.d/admin 'sync'. [production]
14:06 <jmm@cumin2003> START - Cookbook sre.hosts.reboot-single for host ganeti1041.eqiad.wmnet [production]
14:04 <tappof@cumin1003> START - Cookbook sre.hosts.reboot-single for host prometheus1006.eqiad.wmnet [production]
14:03 <tappof@cumin1003> START - Cookbook sre.hosts.reboot-single for host prometheus2007.codfw.wmnet [production]
14:03 <tappof@cumin1003> END (PASS) - Cookbook sre.hosts.reboot-single (exit_code=0) for host prometheus2005.codfw.wmnet [production]
14:03 <tappof@cumin1003> END (PASS) - Cookbook sre.hosts.reboot-single (exit_code=0) for host prometheus1007.eqiad.wmnet [production]
14:03 <jmm@cumin2003> END (PASS) - Cookbook sre.hosts.reboot-single (exit_code=0) for host netflow1002.eqiad.wmnet [production]
14:02 <bking@cumin2003> START - Cookbook sre.hosts.downtime for 2:00:00 on apifeatureusage1001.eqiad.wmnet with reason: host reimage [production]
14:02 <cgoubert@cumin2003> conftool action : set/pooled=true; selector: dnsdisc=helm-charts.*,name=eqiad [production]
14:01 <cgoubert@cumin2003> END (PASS) - Cookbook sre.hosts.reboot-single (exit_code=0) for host chartmuseum1001.eqiad.wmnet [production]
14:00 <moritzm> installing libxml2 security updates [production]
13:59 <btullis@cumin1003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host an-worker1214.eqiad.wmnet with OS bookworm [production]
13:59 <cmooney@cumin1003> conftool action : set/pooled=yes; selector: name=dns5003.* [production]
13:59 <jmm@cumin2003> START - Cookbook sre.ganeti.drain-node for draining ganeti node ganeti1041.eqiad.wmnet [production]
13:59 <jmm@cumin2003> START - Cookbook sre.hosts.reboot-single for host netflow1002.eqiad.wmnet [production]
13:58 <cmooney@dns3003> END - running authdns-update [production]
13:57 <cgoubert@cumin2003> START - Cookbook sre.hosts.reboot-single for host chartmuseum1001.eqiad.wmnet [production]
13:57 <cgoubert@cumin2003> conftool action : set/pooled=false; selector: dnsdisc=helm-charts.*,name=eqiad [production]
13:57 <bking@cumin2003> END (PASS) - Cookbook sre.elasticsearch.rolling-operation (exit_code=0) Operation.RESTART (1 nodes at a time) for ElasticSearch cluster cloudelastic: cloudelastic cluster restart - bking@cumin2003 [production]
13:57 <cgoubert@cumin2003> conftool action : set/pooled=true; selector: dnsdisc=helm-charts.*,name=codfw [production]
13:56 <cmooney@dns3003> START - running authdns-update [production]
13:56 <jayme@deploy1003> helmfile [codfw] START helmfile.d/admin 'sync'. [production]
13:56 <cmooney@cumin1003> conftool action : set/pooled=yes; selector: name=dns5003.*,service=authdns-update [production]
13:55 <cgoubert@cumin2003> END (PASS) - Cookbook sre.hosts.reboot-single (exit_code=0) for host chartmuseum2001.codfw.wmnet [production]
13:55 <tappof@cumin1003> START - Cookbook sre.hosts.reboot-single for host prometheus1007.eqiad.wmnet [production]
13:55 <cmooney@cumin1003> END (PASS) - Cookbook sre.hosts.reboot-single (exit_code=0) for host dns5003.wikimedia.org [production]
13:55 <tappof@cumin1003> END (PASS) - Cookbook sre.hosts.reboot-single (exit_code=0) for host prometheus1005.eqiad.wmnet [production]
13:51 <cgoubert@cumin2003> START - Cookbook sre.hosts.reboot-single for host chartmuseum2001.codfw.wmnet [production]
13:51 <cgoubert@cumin2003> conftool action : set/pooled=false; selector: dnsdisc=helm-charts.*,name=codfw [production]
13:51 <tappof@cumin1003> START - Cookbook sre.hosts.reboot-single for host prometheus2005.codfw.wmnet [production]
13:51 <bking@cumin2003> START - Cookbook sre.hosts.reimage for host apifeatureusage1001.eqiad.wmnet with OS bookworm [production]
13:50 <tappof@cumin1003> END (PASS) - Cookbook sre.hosts.reboot-single (exit_code=0) for host prometheus7002.magru.wmnet [production]