51-100 of 10000 results (121ms)
2026-08-03 ยง
10:46 <cmooney@cumin1003> START - Cookbook sre.puppet.sync-netbox-hiera generate netbox hiera data: "Triggered by cookbooks.sre.dns.netbox: add new reverse ranges for eqsin CR switch links - cmooney@cumin1003" [production]
10:42 <ladsgroup@deploy1003> Started scap sync-world: Backport for [[gerrit:1320115|Remove more $wmg = $wg hacks (T119117)]] [production]
10:41 <cmooney@cumin1003> START - Cookbook sre.dns.netbox [production]
10:36 <marostegui@cumin1003> dbctl commit (dc=all): 'Depool db2245, db2246 and db2247 T433610', diff saved to https://phabricator.wikimedia.org/P95855 and previous config saved to /var/cache/conftool/dbconfig/20260803-103652-marostegui.json [production]
10:35 <marostegui@cumin1003> dbctl commit (dc=all): 'Depool db2248 from s4 T433610', diff saved to https://phabricator.wikimedia.org/P95854 and previous config saved to /var/cache/conftool/dbconfig/20260803-103535-marostegui.json [production]
10:27 <cgoubert@deploy1003> helmfile [eqiad] DONE helmfile.d/services/rest-gateway: apply [production]
10:27 <cgoubert@deploy1003> helmfile [eqiad] START helmfile.d/services/rest-gateway: apply [production]
10:26 <cgoubert@deploy1003> helmfile [codfw] DONE helmfile.d/services/rest-gateway: apply [production]
10:24 <kart_> cxserver: Add referencePunctuation config (T97231) [production]
10:24 <cgoubert@deploy1003> helmfile [codfw] START helmfile.d/services/rest-gateway: apply [production]
10:23 <cgoubert@deploy1003> helmfile [staging] DONE helmfile.d/services/rest-gateway: apply [production]
10:23 <cgoubert@deploy1003> helmfile [staging] START helmfile.d/services/rest-gateway: apply [production]
10:23 <cgoubert@deploy1003> helmfile [staging] DONE helmfile.d/services/rest-gateway: apply [production]
10:22 <cgoubert@deploy1003> helmfile [staging] START helmfile.d/services/rest-gateway: apply [production]
10:22 <kartik@deploy1003> helmfile [eqiad] DONE helmfile.d/services/cxserver: apply [production]
10:21 <kartik@deploy1003> helmfile [eqiad] START helmfile.d/services/cxserver: apply [production]
10:21 <mvernon@cumin2003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host ms-be2085.codfw.wmnet with OS trixie [production]
10:20 <kartik@deploy1003> helmfile [codfw] DONE helmfile.d/services/cxserver: apply [production]
10:20 <kartik@deploy1003> helmfile [codfw] START helmfile.d/services/cxserver: apply [production]
10:18 <kartik@deploy1003> helmfile [staging] DONE helmfile.d/services/cxserver: apply [production]
10:18 <kartik@deploy1003> helmfile [staging] START helmfile.d/services/cxserver: apply [production]
10:01 <mvernon@cumin1003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host ms-be1087.eqiad.wmnet with OS trixie [production]
09:48 <mvernon@cumin2003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on ms-be2085.codfw.wmnet with reason: host reimage [production]
09:43 <mvernon@cumin1003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on ms-be1087.eqiad.wmnet with reason: host reimage [production]
09:40 <mvernon@cumin2003> START - Cookbook sre.hosts.downtime for 2:00:00 on ms-be2085.codfw.wmnet with reason: host reimage [production]
09:40 <mvernon@cumin1003> START - Cookbook sre.hosts.downtime for 2:00:00 on ms-be1087.eqiad.wmnet with reason: host reimage [production]
09:26 <mvernon@cumin1003> START - Cookbook sre.hosts.reimage for host ms-be1087.eqiad.wmnet with OS trixie [production]
09:26 <mvernon@cumin2003> START - Cookbook sre.hosts.reimage for host ms-be2085.codfw.wmnet with OS trixie [production]
09:13 <mvernon@cumin2003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host ms-be2084.codfw.wmnet with OS trixie [production]
09:01 <mvernon@cumin1003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host ms-be1086.eqiad.wmnet with OS trixie [production]
08:54 <mvernon@cumin2003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on ms-be2084.codfw.wmnet with reason: host reimage [production]
08:50 <mvernon@cumin2003> START - Cookbook sre.hosts.downtime for 2:00:00 on ms-be2084.codfw.wmnet with reason: host reimage [production]
08:44 <mvernon@cumin1003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on ms-be1086.eqiad.wmnet with reason: host reimage [production]
08:39 <mvernon@cumin1003> START - Cookbook sre.hosts.downtime for 2:00:00 on ms-be1086.eqiad.wmnet with reason: host reimage [production]
08:38 <arthurtaylor@deploy1003> helmfile [eqiad] DONE helmfile.d/services/wikidata-query-gui: apply [production]
08:38 <arthurtaylor@deploy1003> helmfile [eqiad] START helmfile.d/services/wikidata-query-gui: apply [production]
08:37 <arthurtaylor@deploy1003> helmfile [codfw] DONE helmfile.d/services/wikidata-query-gui: apply [production]
08:37 <arthurtaylor@deploy1003> helmfile [codfw] START helmfile.d/services/wikidata-query-gui: apply [production]
08:35 <mvernon@cumin2003> START - Cookbook sre.hosts.reimage for host ms-be2084.codfw.wmnet with OS trixie [production]
08:34 <arthurtaylor@deploy1003> helmfile [staging] DONE helmfile.d/services/wikidata-query-gui: apply [production]
08:34 <arthurtaylor@deploy1003> helmfile [staging] START helmfile.d/services/wikidata-query-gui: apply [production]
08:27 <ozge@deploy1003> helmfile [ml-serve-eqiad] Ran 'sync' command on namespace 'llm' for release 'main' . [production]
08:25 <mvernon@cumin1003> START - Cookbook sre.hosts.reimage for host ms-be1086.eqiad.wmnet with OS trixie [production]
08:09 <marostegui@cumin1003> END (PASS) - Cookbook sre.mysql.pool (exit_code=0) pool db1218: Repool after a crash [production]
08:07 <mvernon@cumin2003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host ms-be2083.codfw.wmnet with OS trixie [production]
08:06 <mvernon@cumin1003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host ms-be1085.eqiad.wmnet with OS trixie [production]
07:48 <mvernon@cumin2003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on ms-be2083.codfw.wmnet with reason: host reimage [production]
07:44 <mvernon@cumin1003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on ms-be1085.eqiad.wmnet with reason: host reimage [production]
07:40 <kart_> Updated cxsever to 2026-07-16-140518-production (T97231) [production]
07:39 <kartik@deploy1003> helmfile [eqiad] DONE helmfile.d/services/cxserver: apply [production]