151-200 of 10000 results (13ms)
2026-08-13 §
17:41 <bking@cumin2003> START - Cookbook sre.ganeti.makevm for new host dse-k8s-etcd2001.codfw.wmnet [production]
17:40 <inflatador> bking@ganeti2048] `sudo gnt-instance remove --force --ignore-failures --shutdown-timeout=0` on non-DRBD VMs T434681 [production]
16:20 <bking@cumin2003> END (PASS) - Cookbook sre.elasticsearch.rolling-operation (exit_code=0) Operation.RESTART (3 nodes at a time) for ElasticSearch cluster search_codfw: quash java safepoint logspam - bking@cumin2003 - T434685 [production]
15:39 <inflatador> bking@ganeti2048] ~$ sudo gnt-node failover -f --ignore-consistency ganeti2046.codfw.wmnet T434681 [production]
15:18 <inflatador> bking@ganeti2048 sudo gnt-node failover -f ganeti2046.codfw.wmnet T434681 [production]
15:09 <bking@cumin2003> START - Cookbook sre.elasticsearch.rolling-operation Operation.RESTART (3 nodes at a time) for ElasticSearch cluster search_codfw: quash java safepoint logspam - bking@cumin2003 - T434685 [production]
14:59 <bking@cumin2003> END (FAIL) - Cookbook sre.ganeti.makevm (exit_code=99) for new host dse-k8s-etcd2004.codfw.wmnet [production]
14:58 <bking@cumin2003> END (PASS) - Cookbook sre.dns.wipe-cache (exit_code=0) dse-k8s-etcd2004.codfw.wmnet on all recursors [production]
14:58 <bking@cumin2003> START - Cookbook sre.dns.wipe-cache dse-k8s-etcd2004.codfw.wmnet on all recursors [production]
14:58 <bking@cumin2003> END (PASS) - Cookbook sre.dns.netbox (exit_code=0) [production]
14:58 <bking@cumin2003> END (PASS) - Cookbook sre.puppet.sync-netbox-hiera (exit_code=0) generate netbox hiera data: "Triggered by cookbooks.sre.dns.netbox: Remove records for VM dse-k8s-etcd2004.codfw.wmnet - bking@cumin2003" [production]
14:58 <bking@cumin2003> START - Cookbook sre.puppet.sync-netbox-hiera generate netbox hiera data: "Triggered by cookbooks.sre.dns.netbox: Remove records for VM dse-k8s-etcd2004.codfw.wmnet - bking@cumin2003" [production]
14:53 <bking@cumin2003> START - Cookbook sre.dns.netbox [production]
14:53 <bking@cumin2003> END (PASS) - Cookbook sre.dns.wipe-cache (exit_code=0) dse-k8s-etcd2004.codfw.wmnet on all recursors [production]
14:53 <bking@cumin2003> START - Cookbook sre.dns.wipe-cache dse-k8s-etcd2004.codfw.wmnet on all recursors [production]
14:53 <bking@cumin2003> END (PASS) - Cookbook sre.dns.netbox (exit_code=0) [production]
14:53 <bking@cumin2003> END (PASS) - Cookbook sre.puppet.sync-netbox-hiera (exit_code=0) generate netbox hiera data: "Triggered by cookbooks.sre.dns.netbox: Add records for VM dse-k8s-etcd2004.codfw.wmnet - bking@cumin2003" [production]
14:53 <bking@cumin2003> START - Cookbook sre.puppet.sync-netbox-hiera generate netbox hiera data: "Triggered by cookbooks.sre.dns.netbox: Add records for VM dse-k8s-etcd2004.codfw.wmnet - bking@cumin2003" [production]
14:48 <bking@cumin2003> START - Cookbook sre.dns.netbox [production]
14:48 <bking@cumin2003> START - Cookbook sre.ganeti.makevm for new host dse-k8s-etcd2004.codfw.wmnet [production]
14:09 <bking@cumin2003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on apifeatureusage1001.eqiad.wmnet with reason: host reimage [production]
14:02 <bking@cumin2003> START - Cookbook sre.hosts.downtime for 2:00:00 on apifeatureusage1001.eqiad.wmnet with reason: host reimage [production]
13:57 <bking@cumin2003> END (PASS) - Cookbook sre.elasticsearch.rolling-operation (exit_code=0) Operation.RESTART (1 nodes at a time) for ElasticSearch cluster cloudelastic: cloudelastic cluster restart - bking@cumin2003 [production]
13:51 <bking@cumin2003> START - Cookbook sre.hosts.reimage for host apifeatureusage1001.eqiad.wmnet with OS bookworm [production]
13:31 <bking@cumin2003> START - Cookbook sre.elasticsearch.rolling-operation Operation.RESTART (1 nodes at a time) for ElasticSearch cluster cloudelastic: cloudelastic cluster restart - bking@cumin2003 [production]
2026-08-12 §
17:11 <bking@cumin2003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host apifeatureusage2001.codfw.wmnet with OS bookworm [production]
14:09 <bking@cumin2003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on apifeatureusage2001.codfw.wmnet with reason: host reimage [production]
14:02 <bking@cumin2003> START - Cookbook sre.hosts.downtime for 2:00:00 on apifeatureusage2001.codfw.wmnet with reason: host reimage [production]
13:44 <bking@cumin2003> START - Cookbook sre.hosts.reimage for host apifeatureusage2001.codfw.wmnet with OS bookworm [production]
2026-08-05 §
21:12 <bking@cumin2003> conftool action : set/pooled=true; selector: dnsdisc=search-psi,name=eqiad [production]
21:12 <bking@cumin2003> conftool action : set/pooled=true; selector: dnsdisc=search,name=eqiad [production]
21:10 <bking@cumin2003> conftool action : set/pooled=false; selector: dnsdisc=search,name=eqiad [production]
21:07 <bking@cumin2003> conftool action : set/pooled=false; selector: dnsdisc=search-psi,name=eqiad [production]
19:51 <inflatador> [bking@puppetserver1001] ~$ sudo puppetserver ca sign --certname an-worker1189.eqiad.wmnet T434142 [production]
19:47 <bking@cumin2003> DONE (FAIL) - Cookbook sre.puppet.renew-cert (exit_code=99) for an-worker1189.eqiad.wmnet: Renew puppet certificate - bking@cumin2003 [production]
16:55 <bking@cumin2003> conftool action : set/pooled=true; selector: dnsdisc=search-psi,name=eqiad [production]
16:55 <bking@cumin2003> conftool action : set/pooled=true; selector: dnsdisc=search-omega,name=eqiad [production]
16:55 <bking@cumin2003> conftool action : set/pooled=true; selector: dnsdisc=search,name=eqiad [production]
16:53 <bking@cumin2003> END (PASS) - Cookbook sre.elasticsearch.rolling-operation (exit_code=0) Operation.RESTART (3 nodes at a time) for ElasticSearch cluster search_eqiad: apply logging and security config updates - bking@cumin2003 - T324335 [production]
16:40 <bking@cumin2003> START - Cookbook sre.elasticsearch.rolling-operation Operation.RESTART (3 nodes at a time) for ElasticSearch cluster search_eqiad: apply logging and security config updates - bking@cumin2003 - T324335 [production]
16:34 <bking@cumin2003> END (FAIL) - Cookbook sre.elasticsearch.rolling-operation (exit_code=99) Operation.RESTART (3 nodes at a time) for ElasticSearch cluster search_eqiad: apply logging and security config updates - bking@cumin2003 - T324335 [production]
15:38 <bking@cumin2003> START - Cookbook sre.elasticsearch.rolling-operation Operation.RESTART (3 nodes at a time) for ElasticSearch cluster search_eqiad: apply logging and security config updates - bking@cumin2003 - T324335 [production]
15:32 <bking@cumin2003> END (FAIL) - Cookbook sre.elasticsearch.rolling-operation (exit_code=99) Operation.RESTART (3 nodes at a time) for ElasticSearch cluster search_eqiad: apply logging and security config updates - bking@cumin2003 - T324335 [production]
14:22 <bking@cumin2003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host search-loader1002.eqiad.wmnet with OS trixie [production]
14:04 <bking@cumin2003> START - Cookbook sre.elasticsearch.rolling-operation Operation.RESTART (3 nodes at a time) for ElasticSearch cluster search_eqiad: apply logging and security config updates - bking@cumin2003 - T324335 [production]
14:04 <bking@cumin2003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on search-loader1002.eqiad.wmnet with reason: host reimage [production]
14:01 <bking@cumin2003> conftool action : set/pooled=false; selector: dnsdisc=search-psi,name=eqiad [production]
14:01 <bking@cumin2003> conftool action : set/pooled=false; selector: dnsdisc=search-omega,name=eqiad [production]
14:01 <bking@cumin2003> conftool action : set/pooled=false; selector: dnsdisc=search,name=eqiad [production]
13:57 <bking@cumin2003> START - Cookbook sre.hosts.downtime for 2:00:00 on search-loader1002.eqiad.wmnet with reason: host reimage [production]