Fleet changelogs · dev.ecs0.net
rdmsm4x-changelog-20260822-1917-rdmbair15m5-tailscale-repair

rdmbair15m5 is back on the tailnet: tailscale up brought its already-authenticated node from Stopped to Running, and the scheduled (launchd) fleet probe now sees 6/6 hosts, verified via a forced launchd kickstart, not just an interactive run.

Scope

What was wrong

Commands run (on rdmbair15m5, over [email protected])

brew install tailscale        # already installed/up to date (1.102.3) — confirms CLI presence
tailscale up                  # BackendState Stopped -> Running, no re-auth prompt (rc=0)
tailscale status --json       # confirmed: Running, Self.Online true, direct connection

Commands run (on rdmsm4x)

python3 <one-off script>      # reordered rdmbair15m5's entry in data/fleet_hosts.json using
                               # live `tailscale status` as ground truth; every other host's
                               # cache was already correctly ordered (no-op for them)
touch -t 202608220001 ~/dev/data/fleet.json
launchctl kickstart -k gui/501/com.eastcoastscience.devsite

Verification evidence

Files touched

Backup / undo

Outstanding — needs Rich, not blocking

Every fleet host (rdmbair15m5, rdmbair13m5, rdmpw3265m, rdmpw3275m, and rdmsm4x itself) is carrying one stale duplicate Tailscale node: a live <host>-1 entry plus a bare <host> entry that's been offline 1+ day. Harmless (the live node is what's used) but cluttered, and only the account owner can remove another node's registration — the CLI can't. To clean up: https://login.tailscale.com/admin/machines → for each of the 5 stale (non--1, "Last seen" 1d+) rows → ⋯ menu → Delete. Double-check you're deleting the offline row, not the -1 (active) row, before confirming.