Files
sfunkenhauser 8d5a5b70ba Clean up actors when ateom terminates (#1888)
Fixes #1882

When a worker pod's ateom container restarts, every sandbox it was
hosting is lost, but the control plane kept reporting those Actors as
running.

This change:

- **Tracks ateom restarts:** each Worker gets an `epoch`, set by the
syncer from the ateom container's restart count. It can only increase.
- **Stamps each binding:** every Actor assignment records the Worker's
epoch at bind time.
- **Crashes stale Actors:** a new reconciler in ateapi watches for
Workers whose epoch has risen past `status.observed_epoch`. It crashes
the Actors bound during earlier epochs, releases their assignments, then
advances `observed_epoch`.



[ x ] Tests pass
[ x ] Appropriate changes to documentation are included in the PR
2026-09-30 23:01:13 +00:00
..