201-250 of 10000 results (112ms)
2026-07-20 ยง
14:49 <sukhe@cumin1003> END (PASS) - Cookbook sre.loadbalancer.admin (exit_code=0) rebooting A:liberica and not P{lvs7003*} and A:liberica [production]
14:47 <urbanecm@deploy2003> vadymts1, migr, urbanecm: Backport for [[gerrit:1312471|postEdit experiment: enroll control users by same criteria]], [[gerrit:1312423|Modify user groups rights in English Wikiquote (T432557)]] synced to the testservers (see https://wikitech.wikimedia.org/wiki/Mwdebug). Changes can now be verified there. [production]
14:44 <sukhe@cumin1003> END (PASS) - Cookbook sre.dns.admin (exit_code=0) DNS admin: pool magru [reason: BGP issues in lvs7003 resolved after liberica restart, no task ID specified] [production]
14:44 <sukhe@cumin1003> START - Cookbook sre.dns.admin DNS admin: pool magru [reason: BGP issues in lvs7003 resolved after liberica restart, no task ID specified] [production]
14:41 <sukhe@cumin1003> END (PASS) - Cookbook sre.loadbalancer.upgrade (exit_code=0) restart P{lvs7003.magru.wmnet} and A:liberica [production]
14:41 <sukhe@cumin1003> END (PASS) - Cookbook sre.loadbalancer.admin (exit_code=0) pooling P{lvs7003.magru.wmnet} and A:liberica [production]
14:41 <sukhe@cumin1003> START - Cookbook sre.loadbalancer.admin pooling P{lvs7003.magru.wmnet} and A:liberica [production]
14:41 <sukhe@cumin1003> END (PASS) - Cookbook sre.loadbalancer.admin (exit_code=0) depooling P{lvs7003.magru.wmnet} and A:liberica [production]
14:39 <sukhe@cumin1003> START - Cookbook sre.loadbalancer.admin depooling P{lvs7003.magru.wmnet} and A:liberica [production]
14:39 <sukhe@cumin1003> START - Cookbook sre.loadbalancer.upgrade restart P{lvs7003.magru.wmnet} and A:liberica [production]
14:33 <sukhe@cumin1003> END (PASS) - Cookbook sre.dns.admin (exit_code=0) DNS admin: depool magru [reason: no reason specified, no task ID specified] [production]
14:33 <sukhe@cumin1003> START - Cookbook sre.dns.admin DNS admin: depool magru [reason: no reason specified, no task ID specified] [production]
14:31 <urbanecm@deploy2003> Started scap sync-world: Backport for [[gerrit:1312471|postEdit experiment: enroll control users by same criteria]], [[gerrit:1312423|Modify user groups rights in English Wikiquote (T432557)]] [production]
14:24 <jgiannelos@deploy2003> helmfile [codfw] DONE helmfile.d/services/mw-parsoid: apply [production]
14:24 <jgiannelos@deploy2003> helmfile [codfw] START helmfile.d/services/mw-parsoid: apply [production]
14:24 <jgiannelos@deploy2003> helmfile [eqiad] DONE helmfile.d/services/mw-parsoid: apply [production]
14:24 <jgiannelos@deploy2003> helmfile [eqiad] START helmfile.d/services/mw-parsoid: apply [production]
14:23 <bking@cumin2003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on wdqs2019.codfw.wmnet with reason: host reimage [production]
14:22 <bking@cumin2003> START - Cookbook sre.wdqs.data-transfer (T430880, restore data on newly-reimaged host) xfer scholarly_articles from wdqs2017.codfw.wmnet -> wdqs2027.codfw.wmnet, repooling source-only afterwards [production]
14:19 <bking@cumin2003> START - Cookbook sre.hosts.downtime for 2:00:00 on wdqs2019.codfw.wmnet with reason: host reimage [production]
14:19 <bking@cumin2003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host wdqs2027.codfw.wmnet with OS bookworm [production]
14:16 <sukhe@cumin1003> cookbooks.sre.cdn.roll-reboot finished rebooting cp2046.codfw.wmnet [production]
14:16 <sukhe@cumin1003> cookbooks.sre.cdn.roll-reboot finished rebooting cp2045.codfw.wmnet [production]
14:08 <jgiannelos@deploy2003> helmfile [codfw] DONE helmfile.d/services/mw-parsoid: apply [production]
14:08 <jgiannelos@deploy2003> helmfile [codfw] START helmfile.d/services/mw-parsoid: apply [production]
14:08 <jgiannelos@deploy2003> helmfile [eqiad] DONE helmfile.d/services/mw-parsoid: apply [production]
14:08 <jgiannelos@deploy2003> helmfile [eqiad] START helmfile.d/services/mw-parsoid: apply [production]
14:07 <jgiannelos@deploy2003> helmfile [codfw] DONE helmfile.d/services/mw-parsoid: apply [production]
14:06 <jgiannelos@deploy2003> helmfile [codfw] START helmfile.d/services/mw-parsoid: apply [production]
14:06 <jgiannelos@deploy2003> helmfile [eqiad] DONE helmfile.d/services/mw-parsoid: apply [production]
14:06 <bking@cumin2003> START - Cookbook sre.elasticsearch.rolling-operation Operation.REBOOT (1 nodes at a time) for ElasticSearch cluster cloudelastic: apply security updates - bking@cumin2003 - T431826 [production]
14:05 <jgiannelos@deploy2003> helmfile [eqiad] START helmfile.d/services/mw-parsoid: apply [production]
14:00 <bking@cumin2003> END (PASS) - Cookbook sre.hosts.move-vlan (exit_code=0) for host wdqs2019 [production]
14:00 <bking@cumin2003> END (PASS) - Cookbook sre.network.configure-switch-interfaces (exit_code=0) for host wdqs2019 [production]
13:56 <sukhe@cumin1003> END (PASS) - Cookbook sre.hosts.reboot-single (exit_code=0) for host lvs1015.eqiad.wmnet [production]
13:54 <bking@cumin2003> END (PASS) - Cookbook sre.hosts.downtime (exit_code=0) for 2:00:00 on wdqs2027.codfw.wmnet with reason: host reimage [production]
13:54 <mvernon@cumin2003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host ms-be2071.codfw.wmnet with OS trixie [production]
13:51 <sukhe@cumin1003> START - Cookbook sre.loadbalancer.admin rebooting A:liberica and not P{lvs7003*} and A:liberica [production]
13:51 <sukhe@cumin1003> END (ERROR) - Cookbook sre.loadbalancer.admin (exit_code=97) rebooting A:liberica and P{lvs7003*} and A:liberica [production]
13:51 <sukhe@cumin1003> START - Cookbook sre.loadbalancer.admin rebooting A:liberica and P{lvs7003*} and A:liberica [production]
13:51 <sukhe@cumin1003> START - Cookbook sre.hosts.reboot-single for host lvs1015.eqiad.wmnet [production]
13:50 <sukhe@cumin1003> END (PASS) - Cookbook sre.hosts.reboot-single (exit_code=0) for host lvs1014.eqiad.wmnet [production]
13:50 <mvernon@cumin1003> END (PASS) - Cookbook sre.hosts.reimage (exit_code=0) for host ms-be1076.eqiad.wmnet with OS trixie [production]
13:50 <bking@cumin2003> START - Cookbook sre.hosts.downtime for 2:00:00 on wdqs2027.codfw.wmnet with reason: host reimage [production]
13:45 <sukhe@cumin1003> START - Cookbook sre.hosts.reboot-single for host lvs1014.eqiad.wmnet [production]
13:44 <sukhe@cumin1003> END (PASS) - Cookbook sre.hosts.reboot-single (exit_code=0) for host lvs1013.eqiad.wmnet [production]
13:39 <kevinbazira@deploy2003> helmfile [ml-staging-codfw] 'sync' command on namespace 'tts-section-generator' for release 'main' . [production]
13:38 <sukhe@cumin1003> START - Cookbook sre.hosts.reboot-single for host lvs1013.eqiad.wmnet [production]
13:37 <sukhe@cumin1003> cookbooks.sre.cdn.roll-reboot finished rebooting cp2044.codfw.wmnet [production]
13:37 <sukhe@cumin1003> cookbooks.sre.cdn.roll-reboot finished rebooting cp2043.codfw.wmnet [production]