101-150 of 10000 results (21ms)
2026-07-29 §
18:13 <brett@cumin2002> START - Cookbook sre.loadbalancer.restart-pybal rolling-restart of pybal on P{lvs2014.codfw.wmnet} and A:lvs (T428495) [production]
18:08 <bking@cumin2003> END (PASS) - Cookbook sre.hosts.move-vlan (exit_code=0) for host wdqs2022 [production]
18:08 <bking@cumin2003> END (PASS) - Cookbook sre.network.configure-switch-interfaces (exit_code=0) for host wdqs2022 [production]
18:03 <swfrench-wmf> restarted navtiming on webperf2003 - T428495 [production]
18:03 <bking@cumin2003> START - Cookbook sre.network.configure-switch-interfaces for host wdqs2022 [production]
18:02 <bking@cumin2003> END (PASS) - Cookbook sre.dns.wipe-cache (exit_code=0) wdqs2022.codfw.wmnet 211.48.192.10.in-addr.arpa 1.1.2.0.8.4.0.0.2.9.1.0.0.1.0.0.4.0.1.0.0.6.8.0.0.0.0.0.0.2.6.2.ip6.arpa on all recursors [production]
18:02 <bking@cumin2003> START - Cookbook sre.dns.wipe-cache wdqs2022.codfw.wmnet 211.48.192.10.in-addr.arpa 1.1.2.0.8.4.0.0.2.9.1.0.0.1.0.0.4.0.1.0.0.6.8.0.0.0.0.0.0.2.6.2.ip6.arpa on all recursors [production]
18:02 <bking@cumin2003> END (PASS) - Cookbook sre.dns.netbox (exit_code=0) [production]
18:02 <bking@cumin2003> END (PASS) - Cookbook sre.puppet.sync-netbox-hiera (exit_code=0) generate netbox hiera data: "Triggered by cookbooks.sre.dns.netbox: Update records for host wdqs2022 - bking@cumin2003" [production]
18:02 <bking@cumin2003> START - Cookbook sre.puppet.sync-netbox-hiera generate netbox hiera data: "Triggered by cookbooks.sre.dns.netbox: Update records for host wdqs2022 - bking@cumin2003" [production]
17:57 <bking@cumin2003> START - Cookbook sre.dns.netbox [production]
17:56 <brett@cumin2002> END (FAIL) - Cookbook sre.loadbalancer.restart-pybal (exit_code=1) rolling-restart of pybal on A:lvs-codfw and A:lvs (T428495) [production]
17:55 <swfrench-wmf> begin rolling restart of confd in codfw, eqsin, ulsfo - T428495 [production]
17:54 <bking@cumin2003> START - Cookbook sre.hosts.move-vlan for host wdqs2022 [production]
17:50 <bking@cumin2003> START - Cookbook sre.hosts.reimage for host wdqs2022.codfw.wmnet with OS bookworm [production]
17:50 <brett@cumin2002> START - Cookbook sre.loadbalancer.restart-pybal rolling-restart of pybal on A:lvs-codfw and A:lvs (T428495) [production]
17:47 <bking@cumin2003> START - Cookbook sre.wdqs.data-transfer (T430880, restore data on newly-reimaged host) xfer wdqs-all from wdqs2015.codfw.wmnet -> wdqs2021.codfw.wmnet, repooling source-only afterwards [production]
17:47 <bking@deploy1003> Finished deploy [wdqs/wdqs@e8fb00c]: T430880 (duration: 00m 14s) [production]
17:47 <swfrench-wmf> authdns-update to direct codfw, eqsin, ulsfo etcd clients back to codfw - T428495 [production]
17:47 <swfrench@dns1004> END - running authdns-update [production]
17:47 <bking@deploy1003> Started deploy [wdqs/wdqs@e8fb00c]: T430880 [production]
17:47 <cwilliams@cumin1003> dbctl commit (dc=all): 'Depooling db1252 (T431660)', diff saved to https://phabricator.wikimedia.org/P95692 and previous config saved to /var/cache/conftool/dbconfig/20260729-174713-cwilliams.json [production]
17:47 <cwilliams@cumin1003> DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 1 day, 0:00:00 on db1252.eqiad.wmnet with reason: Maintenance [production]
17:46 <root@cumin1003> END (PASS) - Cookbook sre.mysql.pool (exit_code=0) pool db1249: Maintenance [production]
17:45 <bking@cumin2003> conftool action : set/pooled=yes; selector: name=wdqs2015\.codfw\.wmnet,dc=codfw,cluster=wdqs\-main,service=wdqs\-main [production]
17:45 <swfrench@dns1004> START - running authdns-update [production]
17:44 <root@cumin1003> START - Cookbook sre.mysql.pool pool db1202: Maintenance [production]
17:41 <akhatun> Deployed refinery using scap, then deployed onto hdfs [analytics]
17:41 <akhatun> Deployed refinery using scap, then deployed onto hdfs [production]
17:38 <cwilliams@cumin1003> dbctl commit (dc=all): 'Depooling db1202 (T431660)', diff saved to https://phabricator.wikimedia.org/P95688 and previous config saved to /var/cache/conftool/dbconfig/20260729-173759-cwilliams.json [production]
17:37 <cwilliams@cumin1003> DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 1 day, 0:00:00 on db1202.eqiad.wmnet with reason: Maintenance [production]
17:37 <root@cumin1003> END (PASS) - Cookbook sre.mysql.pool (exit_code=0) pool db1194: Maintenance [production]
17:37 <root@cumin1003> START - Cookbook sre.mysql.pool pool db1235: Maintenance [production]
17:31 <root@cumin1003> END (PASS) - Cookbook sre.mysql.pool (exit_code=0) pool db1230: Maintenance [production]
17:30 <cwilliams@cumin1003> dbctl commit (dc=all): 'Depooling db1235 (T431660)', diff saved to https://phabricator.wikimedia.org/P95684 and previous config saved to /var/cache/conftool/dbconfig/20260729-173051-cwilliams.json [production]
17:30 <cwilliams@cumin1003> DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 1 day, 0:00:00 on db1235.eqiad.wmnet with reason: Maintenance [production]
17:30 <root@cumin1003> END (PASS) - Cookbook sre.mysql.pool (exit_code=0) pool db1234: Maintenance [production]
17:26 <akhatun@deploy1003> Finished deploy [analytics/refinery@5669567] (thin): Regular analytics weekly train THIN [analytics/refinery@56695674] (duration: 02m 02s) [production]
17:24 <akhatun@deploy1003> Started deploy [analytics/refinery@5669567] (thin): Regular analytics weekly train THIN [analytics/refinery@56695674] [production]
17:23 <akhatun@deploy1003> Finished deploy [analytics/refinery@5669567]: Regular analytics weekly train [analytics/refinery@56695674] (duration: 06m 20s) [production]
17:20 <dancy@deploy1003> Finished scap sync-world: Testing delay_messageblobstore_purge: true (duration: 06m 29s) [production]
17:17 <akhatun@deploy1003> Started deploy [analytics/refinery@5669567]: Regular analytics weekly train [analytics/refinery@56695674] [production]
17:17 <akhatun@deploy1003> Finished deploy [analytics/refinery@5669567] (hadoop-test): Regular analytics weekly train TEST [analytics/refinery@56695674] (duration: 00m 22s) [production]
17:16 <akhatun@deploy1003> Started deploy [analytics/refinery@5669567] (hadoop-test): Regular analytics weekly train TEST [analytics/refinery@56695674] [production]
17:13 <dancy@deploy1003> Started scap sync-world: Testing delay_messageblobstore_purge: true [production]
17:05 <mutante> CI: contint1002/contint2002 - restarted httpd to be extra sure all is cleaned up - https://integration.wikimedia.org/ci/ is up and running T418521 [production]
17:04 <jhancock@cumin2002> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on ms-be2097.codfw.wmnet with reason: host reimage [production]
17:03 <mutante> CI: contint1002/contint2002 - rm /etc/apache2/jenkins_proxy - removing legacy jenkins proxy config - jenkins is on new dedicated machines and uses jenkins_proxy_ext config T418521 [production]
17:02 <lucaswerkmeister-wmde@deploy1003> Finished scap sync-world: Backport for [[gerrit:1318707|Optimize language name loading with fallbacks (T231755)]] (duration: 36m 25s) [production]
17:00 <root@cumin1003> START - Cookbook sre.mysql.pool pool db1249: Maintenance [production]