1-50 of 10000 results (16ms)
2026-08-07 §
10:06 <marostegui@cumin1003> DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 1 day, 0:00:00 on 21 hosts with reason: cloning [production]
10:01 <marostegui@cumin1003> END (PASS) - Cookbook sre.mysql.depool (exit_code=0) depool db1156: Depool db1156.eqiad.wmnet to then clone it to db1271.eqiad.wmnet - marostegui@cumin1003 [production]
09:59 <marostegui@cumin1003> START - Cookbook sre.mysql.depool depool db1156: Depool db1156.eqiad.wmnet to then clone it to db1271.eqiad.wmnet - marostegui@cumin1003 [production]
09:59 <marostegui@cumin1003> START - Cookbook sre.mysql.clone of db1156.eqiad.wmnet onto db1271.eqiad.wmnet [production]
09:15 <jynus> started stress testing db1245 dbs T431115 [production]
08:19 <bwojtowicz@deploy1003> helmfile [ml-serve-codfw] Ran 'sync' command on namespace 'experimental' for release 'main' . [production]
08:18 <bwojtowicz@deploy1003> helmfile [ml-serve-eqiad] Ran 'sync' command on namespace 'experimental' for release 'main' . [production]
08:16 <bwojtowicz@deploy1003> helmfile [ml-serve-codfw] Ran 'sync' command on namespace 'revscoring-editquality-reverted' for release 'main' . [production]
08:14 <bwojtowicz@deploy1003> helmfile [ml-serve-eqiad] Ran 'sync' command on namespace 'revscoring-editquality-reverted' for release 'main' . [production]
08:13 <bwojtowicz@deploy1003> helmfile [ml-staging-codfw] Ran 'sync' command on namespace 'revscoring-editquality-reverted' for release 'main' . [production]
08:10 <bwojtowicz@deploy1003> helmfile [ml-serve-codfw] Ran 'sync' command on namespace 'revscoring-editquality-goodfaith' for release 'main' . [production]
08:06 <bwojtowicz@deploy1003> helmfile [ml-serve-eqiad] Ran 'sync' command on namespace 'revscoring-editquality-goodfaith' for release 'main' . [production]
08:05 <bwojtowicz@deploy1003> helmfile [ml-staging-codfw] Ran 'sync' command on namespace 'revscoring-editquality-goodfaith' for release 'main' . [production]
08:00 <klausman@cumin1003> DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 10 days, 0:00:00 on ml-serve1015.eqiad.wmnet with reason: Downtime to get full picture of current BIOS settings beyond what Redfish shows [production]
08:00 <bwojtowicz@deploy1003> helmfile [ml-serve-codfw] Ran 'sync' command on namespace 'revscoring-editquality-damaging' for release 'main' . [production]
07:54 <bwojtowicz@deploy1003> helmfile [ml-serve-eqiad] Ran 'sync' command on namespace 'revscoring-editquality-damaging' for release 'main' . [production]
07:54 <bwojtowicz@deploy1003> helmfile [ml-staging-codfw] Ran 'sync' command on namespace 'revscoring-editquality-damaging' for release 'main' . [production]
07:53 <bwojtowicz@deploy1003> helmfile [ml-serve-codfw] Ran 'sync' command on namespace 'revscoring-drafttopic' for release 'main' . [production]
07:52 <bwojtowicz@deploy1003> helmfile [ml-serve-eqiad] Ran 'sync' command on namespace 'revscoring-drafttopic' for release 'main' . [production]
07:51 <bwojtowicz@deploy1003> helmfile [ml-staging-codfw] Ran 'sync' command on namespace 'revscoring-drafttopic' for release 'main' . [production]
07:50 <bwojtowicz@deploy1003> helmfile [ml-serve-codfw] Ran 'sync' command on namespace 'revscoring-draftquality' for release 'main' . [production]
07:49 <bwojtowicz@deploy1003> helmfile [ml-serve-eqiad] Ran 'sync' command on namespace 'revscoring-draftquality' for release 'main' . [production]
07:48 <bwojtowicz@deploy1003> helmfile [ml-staging-codfw] Ran 'sync' command on namespace 'revscoring-draftquality' for release 'main' . [production]
07:47 <bwojtowicz@deploy1003> helmfile [ml-serve-codfw] Ran 'sync' command on namespace 'revscoring-articletopic' for release 'main' . [production]
07:45 <bwojtowicz@deploy1003> helmfile [ml-serve-eqiad] Ran 'sync' command on namespace 'revscoring-articletopic' for release 'main' . [production]
07:45 <bwojtowicz@deploy1003> helmfile [ml-staging-codfw] Ran 'sync' command on namespace 'revscoring-articletopic' for release 'main' . [production]
07:41 <bwojtowicz@deploy1003> helmfile [ml-serve-codfw] Ran 'sync' command on namespace 'revscoring-articlequality' for release 'main' . [production]
07:38 <bwojtowicz@deploy1003> helmfile [ml-serve-eqiad] Ran 'sync' command on namespace 'revscoring-articlequality' for release 'main' . [production]
07:37 <bwojtowicz@deploy1003> helmfile [ml-staging-codfw] Ran 'sync' command on namespace 'revscoring-articlequality' for release 'main' . [production]
07:25 <hashar> Tag Quibble 1.19.0 @ a8a84ed1c0adb34688ce22066ebde8a2d584a7d4 # T432934 T432943 T427922 [releng]
06:35 <jayme> updated istio to 1.29.4 on wikikube eqiad - T427401 [production]
06:08 <marostegui@cumin1003> END (PASS) - Cookbook sre.hosts.decommission (exit_code=0) for hosts db1178.eqiad.wmnet [production]
06:08 <marostegui@cumin1003> END (PASS) - Cookbook sre.dns.netbox (exit_code=0) [production]
06:08 <marostegui@cumin1003> END (PASS) - Cookbook sre.puppet.sync-netbox-hiera (exit_code=0) generate netbox hiera data: "Triggered by cookbooks.sre.dns.netbox: db1178.eqiad.wmnet decommissioned, removing all IPs except the asset tag one - marostegui@cumin1003" [production]
06:06 <marostegui@cumin1003> START - Cookbook sre.puppet.sync-netbox-hiera generate netbox hiera data: "Triggered by cookbooks.sre.dns.netbox: db1178.eqiad.wmnet decommissioned, removing all IPs except the asset tag one - marostegui@cumin1003" [production]
05:55 <marostegui@cumin1003> START - Cookbook sre.dns.netbox [production]
05:49 <marostegui@cumin1003> START - Cookbook sre.hosts.decommission for hosts db1178.eqiad.wmnet [production]
05:46 <marostegui@cumin1003> END (PASS) - Cookbook sre.mysql.decommission (exit_code=0) [production]
05:46 <marostegui@cumin1003> Removing db1178 from zarcillo T433471 [production]
05:45 <marostegui@cumin1003> START - Cookbook sre.mysql.decommission [production]
02:42 <denisse> Extended volume on prometheus2008 for the disk space alert as per https://wikitech.wikimedia.org/wiki/Prometheus#Prometheus_host_running_out_of_space [production]
02:37 <denisse> Extended volume on prometheus2007 tor the disk space alert as per https://wikitech.wikimedia.org/wiki/Prometheus#Prometheus_host_running_out_of_space [production]
02:07 <mwpresync@deploy1003> Finished scap build-images: Publishing wmf/next image (duration: 06m 56s) [production]
02:00 <mwpresync@deploy1003> Started scap build-images: Publishing wmf/next image [production]
2026-08-06 §
21:39 <maryum> Deploy security patch for T433070 [production]
21:29 <maryum> Deploy security patch for T434189 [production]
21:27 <andrew@cloudcumin1001> END (PASS) - Cookbook wmcs.openstack.quota_increase (exit_code=0) by 64 cores, 1600 gigabytes, 4 instances, 131072 ram, 8 volumes [catalyst]
21:27 <andrew@cloudcumin1001> START - Cookbook wmcs.openstack.quota_increase by 64 cores, 1600 gigabytes, 4 instances, 131072 ram, 8 volumes [catalyst]
20:48 <aude@deploy1003> Finished scap sync-world: Backport for [[gerrit:1322038|outreachwiki: disable bureaucrats ability to locally remove users from importer usergroup (T431959)]], [[gerrit:1321639|Turn on feature flag for custom lists for betawiki (T434027)]], [[gerrit:1322054|tcywiki: update logos for 10years anniversary (T434176)]] (duration: 08m 12s) [production]
20:44 <aude@deploy1003> lmora, aude, anzx: Continuing with deployment [production]