1-50 of 10000 results (11ms)
2026-08-03 ยง
17:13 <andrew@cloudcumin1001> START - Cookbook wmcs.toolforge.k8s.reboot for tools-k8s-worker-nfs-67, tools-k8s-worker-nfs-66 (T433426) [tools]
17:13 <andrew@cloudcumin1001> END (FAIL) - Cookbook wmcs.toolforge.k8s.reboot (exit_code=99) for tools-k8s-worker-nfs-67,tools-k8s-worker-nfs-66 (T433426) [tools]
17:13 <andrew@cloudcumin1001> START - Cookbook wmcs.toolforge.k8s.reboot for tools-k8s-worker-nfs-67,tools-k8s-worker-nfs-66 (T433426) [tools]
17:08 <dzahn@cumin2002> START - Cookbook sre.hosts.reimage for host codesearch1001.eqiad.wmnet with OS trixie [production]
17:07 <andrew@cloudcumin1001> END (PASS) - Cookbook wmcs.toolforge.k8s.reboot (exit_code=0) for tools-k8s-worker-nfs-68 (T433426) [tools]
17:07 <andrew@cloudcumin1001> START - Cookbook wmcs.toolforge.k8s.reboot for tools-k8s-worker-nfs-68 (T433426) [tools]
17:06 <dzahn@cumin2002> END (PASS) - Cookbook sre.puppet.sync-netbox-hiera (exit_code=0) generate netbox hiera data: "Triggered by cookbooks.sre.ganeti.makevm: created new VM codesearch1001.eqiad.wmnet - dzahn@cumin2002" [production]
17:06 <rzl@deploy1003> helmfile [codfw] DONE helmfile.d/services/changeprop-jobqueue: apply [production]
17:06 <dzahn@cumin2002> START - Cookbook sre.puppet.sync-netbox-hiera generate netbox hiera data: "Triggered by cookbooks.sre.ganeti.makevm: created new VM codesearch1001.eqiad.wmnet - dzahn@cumin2002" [production]
17:06 <dzahn@cumin2002> END (PASS) - Cookbook sre.dns.wipe-cache (exit_code=0) codesearch1001.eqiad.wmnet on all recursors [production]
17:06 <dzahn@cumin2002> START - Cookbook sre.dns.wipe-cache codesearch1001.eqiad.wmnet on all recursors [production]
17:06 <dzahn@cumin2002> END (PASS) - Cookbook sre.dns.netbox (exit_code=0) [production]
17:06 <dzahn@cumin2002> END (PASS) - Cookbook sre.puppet.sync-netbox-hiera (exit_code=0) generate netbox hiera data: "Triggered by cookbooks.sre.dns.netbox: Add records for VM codesearch1001.eqiad.wmnet - dzahn@cumin2002" [production]
17:05 <andrew@cloudcumin1001> END (ERROR) - Cookbook wmcs.toolforge.k8s.reboot (exit_code=97) for all NFS workers (T433426) [tools]
17:05 <rzl@deploy1003> helmfile [codfw] START helmfile.d/services/changeprop-jobqueue: apply [production]
17:04 <rzl@deploy1003> helmfile [staging] DONE helmfile.d/services/changeprop-jobqueue: apply [production]
17:04 <rzl@deploy1003> helmfile [staging] START helmfile.d/services/changeprop-jobqueue: apply [production]
16:58 <dzahn@cumin2002> START - Cookbook sre.puppet.sync-netbox-hiera generate netbox hiera data: "Triggered by cookbooks.sre.dns.netbox: Add records for VM codesearch1001.eqiad.wmnet - dzahn@cumin2002" [production]
16:57 <mvernon@cumin2003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host ms-be2091.codfw.wmnet with OS trixie [production]
16:54 <ebernhardson@deploy1003> helmfile [staging] DONE helmfile.d/services/cirrus-streaming-updater: apply [production]
16:54 <ebernhardson@deploy1003> helmfile [staging] START helmfile.d/services/cirrus-streaming-updater: apply [production]
16:52 <andrew@cloudcumin1001> END (PASS) - Cookbook wmcs.openstack.cloudvirt.set_maintenance (exit_code=0) (T431374) [admin]
16:52 <andrew@cloudcumin1001> START - Cookbook wmcs.openstack.cloudvirt.set_maintenance (T431374) [admin]
16:49 <ebernhardson@deploy1003> helmfile [codfw] DONE helmfile.d/services/cirrus-streaming-updater: apply [production]
16:49 <ebernhardson@deploy1003> helmfile [codfw] START helmfile.d/services/cirrus-streaming-updater: apply [production]
16:46 <mvernon@cumin1003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host ms-be1092.eqiad.wmnet with OS trixie [production]
16:43 <ebernhardson@deploy1003> helmfile [eqiad] DONE helmfile.d/services/cirrus-streaming-updater: apply [production]
16:43 <ebernhardson@deploy1003> helmfile [eqiad] START helmfile.d/services/cirrus-streaming-updater: apply [production]
16:43 <dzahn@cumin2002> START - Cookbook sre.dns.netbox [production]
16:43 <dzahn@cumin2002> START - Cookbook sre.ganeti.makevm for new host codesearch1001.eqiad.wmnet [production]
16:41 <jhancock@cumin2002> END (PASS) - Cookbook sre.network.configure-switch-interfaces (exit_code=0) for host mc2046 [production]
16:41 <jhancock@cumin2002> START - Cookbook sre.network.configure-switch-interfaces for host mc2046 [production]
16:40 <jhancock@cumin2002> END (PASS) - Cookbook sre.dns.netbox (exit_code=0) [production]
16:40 <hashar> gerrit: added Vaughn Walters to integration group until he get added to the ciadmin LDAP group | T433615 [releng]
16:38 <mvernon@cumin2003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on ms-be2091.codfw.wmnet with reason: host reimage [production]
16:37 <jhancock@cumin2002> START - Cookbook sre.dns.netbox [production]
16:37 <James_F> Zuul: Drop REL1_44 testing, EOL, for T428911 [releng]
16:35 <mvernon@cumin2003> START - Cookbook sre.hosts.downtime for 2:00:00 on ms-be2091.codfw.wmnet with reason: host reimage [production]
16:28 <mvernon@cumin1003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on ms-be1092.eqiad.wmnet with reason: host reimage [production]
16:24 <jiji@deploy1003> helmfile [eqiad] DONE helmfile.d/services/rest-gateway: apply [production]
16:24 <jiji@deploy1003> helmfile [eqiad] START helmfile.d/services/rest-gateway: apply [production]
16:23 <mvernon@cumin1003> START - Cookbook sre.hosts.downtime for 2:00:00 on ms-be1092.eqiad.wmnet with reason: host reimage [production]
16:14 <mvernon@cumin2003> START - Cookbook sre.hosts.reimage for host ms-be2091.codfw.wmnet with OS trixie [production]
16:03 <mvernon@cumin1003> START - Cookbook sre.hosts.reimage for host ms-be1092.eqiad.wmnet with OS trixie [production]
16:02 <mvernon@cumin2003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host ms-be2090.codfw.wmnet with OS trixie [production]
15:51 <jiji@deploy1003> helmfile [codfw] DONE helmfile.d/services/rest-gateway: apply [production]
15:51 <jiji@deploy1003> helmfile [codfw] START helmfile.d/services/rest-gateway: apply [production]
15:51 <jiji@deploy1003> helmfile [staging] DONE helmfile.d/services/rest-gateway: apply [production]
15:50 <jiji@deploy1003> helmfile [staging] START helmfile.d/services/rest-gateway: apply [production]
15:43 <mvernon@cumin2003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on ms-be2090.codfw.wmnet with reason: host reimage [production]