[08:44:22] marostegui, bjensen, today is magru router upgrade day, going to depool the site shortly - https://phabricator.wikimedia.org/T431750 [08:44:49] thanks [08:44:51] sounds good, thanks for the heads up :) [08:44:53] good luck! [08:53:31] btullis: can I deploy Ben Tullis: hadoop-test: Switch to using the new namenodes (818afe156d) ? [08:53:36] Yes please. [08:53:39] cool [08:53:58] Thx [09:02:30] hey folks, I'd need to reboot the idp and idm active hosts, my idea is to reboot them without failover to figure out what is the impact / annoyance if we do it, since in theory the downtime should be a couple of minutes [09:03:04] * marostegui holds to his oncall phone [09:07:20] rebooting now [09:07:28] (idp) [09:08:46] the host is already up [09:11:52] and idm1001 now [09:14:12] and done [09:15:06] nice! no impact at all? [09:20:50] yeah the reboots are super quick, so the failover is overkill imho [09:21:03] sometimes it is good to do it to excercise the procedure etc.. [09:23:57] absolutely, great stuff [10:07:44] all done with magru, giving it a few min then will repool [10:14:14] repooling magru [11:32:43] I am going to deploy proton soon via https://gerrit.wikimedia.org/r/c/operations/deployment-charts/+/1319068, the new image picks up security updates for chromium etc.. [12:19:16] and done [13:04:05] FYI, I plan to pick up the remaining conf* host work in codfw shortly. I anticipate smoother sailing today, but there could of course be surprises lurking. cc: oncallers federico3, Raine [13:04:38] ack thanks