51-100 of 10000 results (23ms)
2026-08-21 ยง
17:12 <lerickson@deploy1003> helmfile [dse-k8s-codfw] START helmfile.d/services/wdqs: apply [production]
17:10 <sukhe@dns1004> END - running authdns-update [production]
17:08 <sukhe@dns1004> START - running authdns-update [production]
17:08 <sukhe@puppetserver1001> conftool action : set/pooled=no; selector: name=dns5004.wikimedia.org [reason: resolving authdns-update issues] [production]
17:07 <sukhe@dns1004> FAIL - running authdns-update [production]
17:05 <lerickson@deploy1003> helmfile [dse-k8s-eqiad] DONE helmfile.d/services/wdqs: apply [production]
17:05 <sukhe@dns1004> START - running authdns-update [production]
17:01 <lerickson@deploy1003> helmfile [dse-k8s-eqiad] START helmfile.d/services/wdqs: apply [production]
16:59 <lerickson@deploy1003> helmfile [dse-k8s-codfw] DONE helmfile.d/services/wdqs: apply [production]
16:56 <lerickson@deploy1003> helmfile [dse-k8s-codfw] START helmfile.d/services/wdqs: apply [production]
16:53 <cdobbins@cumin1003> conftool action : set/pooled=yes; selector: name=dns5004.* [reason: trixie upgrade] [production]
16:52 <cdobbins@cumin1003> END (PASS) - Cookbook sre.hosts.remove-downtime (exit_code=0) for dns5004.wikimedia.org [production]
16:52 <cdobbins@cumin1003> START - Cookbook sre.hosts.remove-downtime for dns5004.wikimedia.org [production]
16:44 <cmooney@dns3003> END - running authdns-update [production]
16:41 <cmooney@dns3003> START - running authdns-update [production]
16:41 <cmooney@cumin1003> END (PASS) - Cookbook sre.dns.netbox (exit_code=0) [production]
16:41 <cmooney@cumin1003> END (PASS) - Cookbook sre.puppet.sync-netbox-hiera (exit_code=0) generate netbox hiera data: "Triggered by cookbooks.sre.dns.netbox: remove dns entries for IPs formerly used on eqsin<->codfw arelion - cmooney@cumin1003" [production]
16:37 <cmooney@cumin1003> START - Cookbook sre.puppet.sync-netbox-hiera generate netbox hiera data: "Triggered by cookbooks.sre.dns.netbox: remove dns entries for IPs formerly used on eqsin<->codfw arelion - cmooney@cumin1003" [production]
16:33 <cmooney@cumin1003> START - Cookbook sre.dns.netbox [production]
16:11 <lerickson@deploy1003> helmfile [dse-k8s-eqiad] DONE helmfile.d/services/wdqs: apply [production]
16:11 <andrew@cloudcumin1001> START - Cookbook wmcs.ceph.osd.bootstrap_and_add [admin]
16:08 <sukhe@dns1004> END - running authdns-update [production]
16:08 <andrew@cloudcumin1001> END (PASS) - Cookbook wmcs.ceph.osd.bootstrap_and_add (exit_code=0) [admin]
16:08 <andrew@cloudcumin1001> START - Cookbook wmcs.ceph.osd.bootstrap_and_add [admin]
16:08 <lerickson@deploy1003> helmfile [dse-k8s-eqiad] START helmfile.d/services/wdqs: apply [production]
16:06 <sukhe@dns1004> START - running authdns-update [production]
16:05 <andrew@cloudcumin1001> END (FAIL) - Cookbook wmcs.ceph.osd.depool_and_destroy (exit_code=99) [admin]
16:05 <andrew@cloudcumin1001> START - Cookbook wmcs.ceph.osd.depool_and_destroy [admin]
16:04 <cmooney@dns3003> END - running authdns-update [production]
16:03 <andrew@cloudcumin1001> END (PASS) - Cookbook wmcs.ceph.osd.bootstrap_and_add (exit_code=0) [admin]
16:03 <andrew@cloudcumin1001> START - Cookbook wmcs.ceph.osd.bootstrap_and_add [admin]
16:02 <cmooney@dns3003> START - running authdns-update [production]
16:00 <cmooney@cumin1003> END (PASS) - Cookbook sre.dns.netbox (exit_code=0) [production]
16:00 <cmooney@cumin1003> END (PASS) - Cookbook sre.puppet.sync-netbox-hiera (exit_code=0) generate netbox hiera data: "Triggered by cookbooks.sre.dns.netbox: remove dns entries for IPs formerly used on eqord<->codfw arelion - cmooney@cumin1003" [production]
15:56 <cdobbins@cumin1003> END (FAIL) - Cookbook sre.hosts.reimage (exit_code=1) for host dns5004.wikimedia.org with OS trixie [production]
15:55 <cmooney@cumin1003> START - Cookbook sre.puppet.sync-netbox-hiera generate netbox hiera data: "Triggered by cookbooks.sre.dns.netbox: remove dns entries for IPs formerly used on eqord<->codfw arelion - cmooney@cumin1003" [production]
15:53 <cmooney@cumin1003> START - Cookbook sre.dns.netbox [production]
15:51 <cmooney@cumin1003> END (ERROR) - Cookbook sre.dns.netbox (exit_code=97) [production]
15:51 <cmooney@cumin1003> START - Cookbook sre.dns.netbox [production]
15:38 <andrew@cumin2003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host cloudcephosd1042.eqiad.wmnet with OS bookworm [production]
15:18 <andrew@cumin2003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on cloudcephosd1042.eqiad.wmnet with reason: host reimage [production]
15:17 <dancy@deploy1003> Finished deploy [gerrit/gerrit@2cc11cc]: Deploying https://gerrit.wikimedia.org/r/c/operations/software/gerrit/+/1327669 (T434726) (duration: 00m 14s) [production]
15:17 <dancy@deploy1003> Started deploy [gerrit/gerrit@2cc11cc]: Deploying https://gerrit.wikimedia.org/r/c/operations/software/gerrit/+/1327669 (T434726) [production]
15:13 <andrew@cumin2003> START - Cookbook sre.hosts.downtime for 2:00:00 on cloudcephosd1042.eqiad.wmnet with reason: host reimage [production]
15:09 <cdobbins@cumin1003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on dns5004.wikimedia.org with reason: host reimage [production]
15:05 <cdobbins@cumin1003> START - Cookbook sre.hosts.downtime for 2:00:00 on dns5004.wikimedia.org with reason: host reimage [production]
14:55 <andrewbogott> reimaging cloudvirt1042, getting a 100% fresh start for T429387 testing [admin]
14:53 <andrew@cumin2003> START - Cookbook sre.hosts.reimage for host cloudcephosd1042.eqiad.wmnet with OS bookworm [production]
14:30 <cdobbins@cumin1003> START - Cookbook sre.hosts.reimage for host dns5004.wikimedia.org with OS trixie [production]
14:29 <cdobbins@cumin1003> conftool action : set/pooled=no; selector: name=dns5004.* [reason: trixie upgrade] [production]