fix(backup-drill): grow drill VM boot volume 20→40 GB (restore ENOSPC) #130
Inga granskare
Etiketter
Inga etiketter
area/backups
area/ci
area/control-panel
area/identity
area/infra
area/observability
area/payments
area/security
area/storage
area/web
blocked
needs-info
needs-triage
ready-for-implementation
type
bug
type
chore
type
docs
type
epic
type
feature
type
task
wontfix
Ingen milstolpe
Inget projekt
Inga tilldelade
1 deltagare
Notiser
Förfallodatum
Inget förfallodatum satt.
Beroenden
Inga beroenden satta
Referens
bitborg/bitborg-infra!130
Läser in…
Hänvisa till i nytt ärende
Ingen beskrivning angiven.
Ta bort grenen "fix/backup-drill-boot-volume-size"
Borttagning av en gren är permanent. Även om den borttagna grenen kan fortsätta existera en kort tid innan den faktiskt tas bort, kan det INTE ångras i de flesta fall. Vill du fortsätta?
The weekly restore drill has been failing (
BackupDrillFailedcritical, firing now). Root cause from the drill logs:The throwaway drill VM boots on a 20 GB Cinder volume (
backup_drill_os_boot_volume_size), but the restored dataset now unpacks to ~25 GB in container storage (the live/srv/gitborg-datavolume is at ~24.6 GB and growing). So the restore runs out of space beforeforgejo doctor— meaning there is currently no verified proof backups can be restored, right before onboarding.Bumps the boot volume to 40 GB (headroom above the live data volume) and rewrites the now-false "20 GiB is enough" comment to explain the sizing rule. The volume is a throwaway scratch disk on a VM that's created and destroyed per drill — no data-loss risk, no prod-service impact; it takes effect on the next drill run after the role is applied.
Validated:
ansible-lint roles/backup-drill/passes (production profile),site.yml --syntax-checkOK. Found by the OpenTofu infra audit (2026-07-19).