51-100 of 10000 results (122ms)
2026-08-20 ยง
15:07 <fceratto@cumin1003> START - Cookbook sre.hosts.downtime for 2:00:00 on db1903.eqiad.wmnet with reason: host reimage [production]
14:54 <fceratto@cumin1003> START - Cookbook sre.hosts.reimage for host db1903.eqiad.wmnet with OS trixie [production]
14:53 <fceratto@cumin1003> END (PASS) - Cookbook sre.puppet.sync-netbox-hiera (exit_code=0) generate netbox hiera data: "Triggered by cookbooks.sre.ganeti.makevm: created new VM db1903.eqiad.wmnet - fceratto@cumin1003" [production]
14:53 <fceratto@cumin1003> START - Cookbook sre.puppet.sync-netbox-hiera generate netbox hiera data: "Triggered by cookbooks.sre.ganeti.makevm: created new VM db1903.eqiad.wmnet - fceratto@cumin1003" [production]
14:53 <fceratto@cumin1003> END (PASS) - Cookbook sre.dns.wipe-cache (exit_code=0) db1903.eqiad.wmnet on all recursors [production]
14:53 <fceratto@cumin1003> START - Cookbook sre.dns.wipe-cache db1903.eqiad.wmnet on all recursors [production]
14:53 <fceratto@cumin1003> END (PASS) - Cookbook sre.dns.netbox (exit_code=0) [production]
14:53 <fceratto@cumin1003> END (PASS) - Cookbook sre.puppet.sync-netbox-hiera (exit_code=0) generate netbox hiera data: "Triggered by cookbooks.sre.dns.netbox: Add records for VM db1903.eqiad.wmnet - fceratto@cumin1003" [production]
14:53 <fceratto@cumin1003> START - Cookbook sre.puppet.sync-netbox-hiera generate netbox hiera data: "Triggered by cookbooks.sre.dns.netbox: Add records for VM db1903.eqiad.wmnet - fceratto@cumin1003" [production]
14:49 <fceratto@cumin1003> START - Cookbook sre.dns.netbox [production]
14:49 <fceratto@cumin1003> START - Cookbook sre.ganeti.makevm for new host db1903.eqiad.wmnet [production]
14:33 <cdobbins@cumin1003> conftool action : set/pooled=yes; selector: name=cp1100.* [production]
14:27 <topranks> reconfigure eqiad<->codfw bgp settings [production]
14:22 <cmooney@cumin1003> END (PASS) - Cookbook sre.dns.netbox (exit_code=0) [production]
14:22 <cmooney@cumin1003> END (PASS) - Cookbook sre.puppet.sync-netbox-hiera (exit_code=0) generate netbox hiera data: "Triggered by cookbooks.sre.dns.netbox: update entries used on new transport backup eqiad codfw - cmooney@cumin1003" [production]
14:19 <cmooney@cumin1003> START - Cookbook sre.puppet.sync-netbox-hiera generate netbox hiera data: "Triggered by cookbooks.sre.dns.netbox: update entries used on new transport backup eqiad codfw - cmooney@cumin1003" [production]
14:14 <bking@deploy1003> helmfile [dse-k8s-eqiad] DONE helmfile.d/dse-k8s-services/opensearch-semantic-search: apply [production]
14:14 <bking@deploy1003> helmfile [dse-k8s-eqiad] START helmfile.d/dse-k8s-services/opensearch-semantic-search: apply [production]
14:13 <moritzm> installing util-linux security updates [production]
14:13 <klausman@cumin1003> END (PASS) - Cookbook sre.k8s.reboot-nodes (exit_code=0) rolling reboot on A:ml-staging-worker [production]
14:13 <klausman@cumin1003> END (PASS) - Cookbook sre.k8s.pool-depool-node (exit_code=0) pool for host ml-staging2003.codfw.wmnet [production]
14:13 <klausman@cumin1003> START - Cookbook sre.k8s.pool-depool-node pool for host ml-staging2003.codfw.wmnet [production]
14:08 <klausman@cumin1003> END (PASS) - Cookbook sre.k8s.pool-depool-node (exit_code=0) depool for host ml-staging2003.codfw.wmnet [production]
14:06 <moritzm> installing libheif security updates [production]
13:58 <klausman@cumin1003> START - Cookbook sre.k8s.pool-depool-node depool for host ml-staging2003.codfw.wmnet [production]
13:58 <klausman@cumin1003> END (PASS) - Cookbook sre.k8s.pool-depool-node (exit_code=0) pool for host ml-staging2002.codfw.wmnet [production]
13:58 <klausman@cumin1003> START - Cookbook sre.k8s.pool-depool-node pool for host ml-staging2002.codfw.wmnet [production]
13:56 <cmooney@cumin1003> START - Cookbook sre.dns.netbox [production]
13:56 <fnegri@cumin1003> END (PASS) - Cookbook sre.hosts.remove-downtime (exit_code=0) for clouddb1025.eqiad.wmnet [production]
13:56 <fnegri@cumin1003> START - Cookbook sre.hosts.remove-downtime for clouddb1025.eqiad.wmnet [production]
13:56 <Lucas_WMDE> UTC afternoon backport+config window done [production]
13:53 <moritzm> installing apr-util security updates [production]
13:51 <klausman@cumin1003> END (PASS) - Cookbook sre.k8s.pool-depool-node (exit_code=0) depool for host ml-staging2002.codfw.wmnet [production]
13:50 <fnegri@cumin1003> conftool action : set/weight=100; selector: name=clouddb1025.eqiad.wmnet [production]
13:49 <fnegri@cumin1003> conftool action : set/pooled=yes; selector: name=clouddb1025.eqiad.wmnet [production]
13:41 <klausman@cumin1003> START - Cookbook sre.k8s.pool-depool-node depool for host ml-staging2002.codfw.wmnet [production]
13:41 <klausman@cumin1003> END (PASS) - Cookbook sre.k8s.pool-depool-node (exit_code=0) pool for host ml-staging2001.codfw.wmnet [production]
13:41 <klausman@cumin1003> START - Cookbook sre.k8s.pool-depool-node pool for host ml-staging2001.codfw.wmnet [production]
13:41 <dpogorzelski@deploy1003> helmfile [ml-staging-codfw] DONE helmfile.d/admin 'sync'. [production]
13:38 <dpogorzelski@deploy1003> helmfile [ml-staging-codfw] START helmfile.d/admin 'sync'. [production]
13:38 <fnegri@cumin1003> DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 1:00:00 on clouddb1025.eqiad.wmnet with reason: Removing s6 from clouddb1025 [production]
13:34 <klausman@cumin1003> END (PASS) - Cookbook sre.k8s.pool-depool-node (exit_code=0) depool for host ml-staging2001.codfw.wmnet [production]
13:31 <dpogorzelski@deploy1003> helmfile [ml-staging-codfw] DONE helmfile.d/admin 'sync'. [production]
13:29 <dpogorzelski@deploy1003> helmfile [ml-staging-codfw] START helmfile.d/admin 'sync'. [production]
13:28 <fnegri@cumin1003> conftool action : set/pooled=no; selector: name=clouddb1025.eqiad.wmnet [production]
13:26 <btullis@cumin1003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host stat1009.eqiad.wmnet with OS bookworm [production]
13:24 <klausman@cumin1003> START - Cookbook sre.k8s.pool-depool-node depool for host ml-staging2001.codfw.wmnet [production]
13:24 <klausman@cumin1003> START - Cookbook sre.k8s.reboot-nodes rolling reboot on A:ml-staging-worker [production]
13:23 <klausman@cumin1003> END (PASS) - Cookbook sre.ganeti.reboot-vm (exit_code=0) for VM ml-serve-ctrl1002.eqiad.wmnet [production]
13:21 <btullis@cumin1003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host stat1010.eqiad.wmnet with OS bookworm [production]