|
2026-07-29
§
|
| 17:47 |
<bking@deploy1003> |
Finished deploy [wdqs/wdqs@e8fb00c]: T430880 (duration: 00m 14s) |
[production] |
| 17:47 |
<swfrench-wmf> |
authdns-update to direct codfw, eqsin, ulsfo etcd clients back to codfw - T428495 |
[production] |
| 17:47 |
<swfrench@dns1004> |
END - running authdns-update |
[production] |
| 17:47 |
<bking@deploy1003> |
Started deploy [wdqs/wdqs@e8fb00c]: T430880 |
[production] |
| 17:47 |
<cwilliams@cumin1003> |
dbctl commit (dc=all): 'Depooling db1252 (T431660)', diff saved to https://phabricator.wikimedia.org/P95692 and previous config saved to /var/cache/conftool/dbconfig/20260729-174713-cwilliams.json |
[production] |
| 17:47 |
<cwilliams@cumin1003> |
DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 1 day, 0:00:00 on db1252.eqiad.wmnet with reason: Maintenance |
[production] |
| 17:46 |
<root@cumin1003> |
END (PASS) - Cookbook sre.mysql.pool (exit_code=0) pool db1249: Maintenance |
[production] |
| 17:45 |
<bking@cumin2003> |
conftool action : set/pooled=yes; selector: name=wdqs2015\.codfw\.wmnet,dc=codfw,cluster=wdqs\-main,service=wdqs\-main |
[production] |
| 17:45 |
<swfrench@dns1004> |
START - running authdns-update |
[production] |
| 17:44 |
<root@cumin1003> |
START - Cookbook sre.mysql.pool pool db1202: Maintenance |
[production] |
| 17:41 |
<akhatun> |
Deployed refinery using scap, then deployed onto hdfs |
[production] |
| 17:38 |
<cwilliams@cumin1003> |
dbctl commit (dc=all): 'Depooling db1202 (T431660)', diff saved to https://phabricator.wikimedia.org/P95688 and previous config saved to /var/cache/conftool/dbconfig/20260729-173759-cwilliams.json |
[production] |
| 17:37 |
<cwilliams@cumin1003> |
DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 1 day, 0:00:00 on db1202.eqiad.wmnet with reason: Maintenance |
[production] |
| 17:37 |
<root@cumin1003> |
END (PASS) - Cookbook sre.mysql.pool (exit_code=0) pool db1194: Maintenance |
[production] |
| 17:37 |
<root@cumin1003> |
START - Cookbook sre.mysql.pool pool db1235: Maintenance |
[production] |
| 17:31 |
<root@cumin1003> |
END (PASS) - Cookbook sre.mysql.pool (exit_code=0) pool db1230: Maintenance |
[production] |
| 17:30 |
<cwilliams@cumin1003> |
dbctl commit (dc=all): 'Depooling db1235 (T431660)', diff saved to https://phabricator.wikimedia.org/P95684 and previous config saved to /var/cache/conftool/dbconfig/20260729-173051-cwilliams.json |
[production] |
| 17:30 |
<cwilliams@cumin1003> |
DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 1 day, 0:00:00 on db1235.eqiad.wmnet with reason: Maintenance |
[production] |
| 17:30 |
<root@cumin1003> |
END (PASS) - Cookbook sre.mysql.pool (exit_code=0) pool db1234: Maintenance |
[production] |
| 17:26 |
<akhatun@deploy1003> |
Finished deploy [analytics/refinery@5669567] (thin): Regular analytics weekly train THIN [analytics/refinery@56695674] (duration: 02m 02s) |
[production] |
| 17:24 |
<akhatun@deploy1003> |
Started deploy [analytics/refinery@5669567] (thin): Regular analytics weekly train THIN [analytics/refinery@56695674] |
[production] |
| 17:23 |
<akhatun@deploy1003> |
Finished deploy [analytics/refinery@5669567]: Regular analytics weekly train [analytics/refinery@56695674] (duration: 06m 20s) |
[production] |
| 17:20 |
<dancy@deploy1003> |
Finished scap sync-world: Testing delay_messageblobstore_purge: true (duration: 06m 29s) |
[production] |
| 17:17 |
<akhatun@deploy1003> |
Started deploy [analytics/refinery@5669567]: Regular analytics weekly train [analytics/refinery@56695674] |
[production] |
| 17:17 |
<akhatun@deploy1003> |
Finished deploy [analytics/refinery@5669567] (hadoop-test): Regular analytics weekly train TEST [analytics/refinery@56695674] (duration: 00m 22s) |
[production] |
| 17:16 |
<akhatun@deploy1003> |
Started deploy [analytics/refinery@5669567] (hadoop-test): Regular analytics weekly train TEST [analytics/refinery@56695674] |
[production] |
| 17:13 |
<dancy@deploy1003> |
Started scap sync-world: Testing delay_messageblobstore_purge: true |
[production] |
| 17:05 |
<mutante> |
CI: contint1002/contint2002 - restarted httpd to be extra sure all is cleaned up - https://integration.wikimedia.org/ci/ is up and running T418521 |
[production] |
| 17:04 |
<jhancock@cumin2002> |
END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on ms-be2097.codfw.wmnet with reason: host reimage |
[production] |
| 17:03 |
<mutante> |
CI: contint1002/contint2002 - rm /etc/apache2/jenkins_proxy - removing legacy jenkins proxy config - jenkins is on new dedicated machines and uses jenkins_proxy_ext config T418521 |
[production] |
| 17:02 |
<lucaswerkmeister-wmde@deploy1003> |
Finished scap sync-world: Backport for [[gerrit:1318707|Optimize language name loading with fallbacks (T231755)]] (duration: 36m 25s) |
[production] |
| 17:00 |
<root@cumin1003> |
START - Cookbook sre.mysql.pool pool db1249: Maintenance |
[production] |
| 16:59 |
<jhancock@cumin2002> |
START - Cookbook sre.hosts.downtime for 2:00:00 on ms-be2097.codfw.wmnet with reason: host reimage |
[production] |
| 16:54 |
<bking@cumin2003> |
END (PASS) - Cookbook sre.wdqs.data-transfer (exit_code=0) (T430880, restore data on newly-reimaged host) xfer wdqs-all from wdqs2020.codfw.wmnet -> wdqs2015.codfw.wmnet, repooling source-only afterwards |
[production] |
| 16:53 |
<cwilliams@cumin1003> |
dbctl commit (dc=all): 'Depooling db1249 (T431660)', diff saved to https://phabricator.wikimedia.org/P95674 and previous config saved to /var/cache/conftool/dbconfig/20260729-165339-cwilliams.json |
[production] |
| 16:53 |
<cwilliams@cumin1003> |
DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 1 day, 0:00:00 on db1249.eqiad.wmnet with reason: Maintenance |
[production] |
| 16:53 |
<root@cumin1003> |
END (PASS) - Cookbook sre.mysql.pool (exit_code=0) pool db1248: Maintenance |
[production] |
| 16:50 |
<root@cumin1003> |
START - Cookbook sre.mysql.pool pool db1194: Maintenance |
[production] |
| 16:47 |
<swfrench-wmf> |
silenced EtcdReplicationDown 57b2b421-1cc9-4e38-9276-94f223fd231c - T428495 |
[production] |
| 16:46 |
<tchin@deploy1003> |
helmfile [dse-k8s-eqiad] DONE helmfile.d/dse-k8s-services/eventstreams-internal: apply |
[production] |
| 16:46 |
<jhancock@cumin2002> |
START - Cookbook sre.hosts.reimage for host ms-be2097.codfw.wmnet with OS bullseye |
[production] |
| 16:46 |
<tchin@deploy1003> |
helmfile [dse-k8s-eqiad] START helmfile.d/dse-k8s-services/eventstreams-internal: apply |
[production] |
| 16:45 |
<root@cumin1003> |
START - Cookbook sre.mysql.pool pool db1230: Maintenance |
[production] |
| 16:44 |
<cwilliams@cumin1003> |
dbctl commit (dc=all): 'Depooling db1194 (T431660)', diff saved to https://phabricator.wikimedia.org/P95669 and previous config saved to /var/cache/conftool/dbconfig/20260729-164422-cwilliams.json |
[production] |
| 16:44 |
<tchin@deploy1003> |
helmfile [eqiad] DONE helmfile.d/services/eventstreams: apply |
[production] |
| 16:44 |
<cwilliams@cumin1003> |
DONE (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 1 day, 0:00:00 on db1194.eqiad.wmnet with reason: Maintenance |
[production] |
| 16:43 |
<root@cumin1003> |
END (PASS) - Cookbook sre.mysql.pool (exit_code=0) pool db1191: Maintenance |
[production] |
| 16:43 |
<btullis@cumin1003> |
START - Cookbook sre.hosts.reboot-single for host an-test-master1003.eqiad.wmnet |
[production] |
| 16:43 |
<root@cumin1003> |
START - Cookbook sre.mysql.pool pool db1234: Maintenance |
[production] |
| 16:43 |
<tchin@deploy1003> |
helmfile [eqiad] START helmfile.d/services/eventstreams: apply |
[production] |