dorfteich/packages/shared/src/system.ts
Claude Fable 5 5cef359b8f
All checks were successful
CI / Lint, typecheck, test (push) Successful in 3m45s
CD / Build and push images (push) Successful in 3m49s
CI / Build container images (push) Has been skipped
CD / Deploy to Test (push) Successful in 9s
CD / Smoke tests against Test (push) Successful in 1m18s
CD / Promote to Int (push) Successful in 11s
CI / Auth e2e pack (push) Successful in 5m35s
CI / Import/export fidelity gate (push) Successful in 47s
Nextcloud backup target: admin-configured, manual + scheduled uploads, in-app restore (#103)
Off-host backups for every self-hoster, configured entirely in the admin
UI — supersedes the host-specific mirror plan behind #84.

shared:
- webdav.ts (new package entry like token-crypto): minimal WebDAV client
  with basic auth — PROPFIND (tolerant multistatus parser), MKCOL, PUT
  (streamed), GET, DELETE; Nextcloud DAV path derived from the plain
  server URL, explicit DAV bases pass through
- backup-status.ts: additive remote-upload status in status.json, the
  restore-status.json contract (running/succeeded/failed + staleness
  bound), the backup_command/backup_maintenance NOTIFY channels, and the
  one-bundle-per-set naming (dorfteich-backup-<id>.tar.gz)
- backup-set.ts moved here from apps/backup (api lists local sets)

backup sidecar:
- reads the backup.* instance settings directly from the database (admin
  changes apply next run; local retention row overrides the env) and the
  app password from the secret store
- after each successful set: bundle dump + files archive + manifest into
  ONE self-contained tar.gz, upload via WebDAV per schedule
  (off/daily/weekly; manual runs always upload), prune remote bundles —
  never the newest — and record the outcome in status.json; upload
  failures alert via a new backupUploadFailed mail (de+en)
- command listener on backup_command (run / restore) with a serial queue
  against the nightly timer
- restore orchestrator: restore-status.json → maintenance NOTIFY →
  grace → (remote: download + manifest-verify bundle) → terminate other
  DB connections → shared perform-restore path (same code as restore.sh)
  → final status + maintenance exit

api:
- MaintenanceGuard (global, registered before the setup gate): 503
  maintenance_mode while restore-status says running; health endpoints
  and the new public GET /backup/restore-status stay exempt; a stale
  running state (crashed sidecar) unblocks after 30 min
- MaintenanceStateService watches the file and restarts the api after a
  successful restore (fresh caches, migrate-on-start for older dumps);
  main.ts refuses to touch the database while a restore runs — a
  container restarting mid-restore must not race pg_restore with
  migrate deploy
- worker sweeps (conversion, mail outbox, scheduler) catch transient
  database failures instead of dying on an unhandled rejection — the
  restore's connection termination crashed the api in verification
- backup admin endpoints under /admin/system/backup: settings (live
  connection test before save, password write-only into the secret
  store), nextcloud/test, sets (local via the ro backups mount + remote
  via WebDAV), run + restore (type-to-confirm backstop, source
  validation) — commands travel as NOTIFY payloads; audit actions
  backup.settings_changed/run_triggered/restore_requested
- readyz: new warning-level backup_remote check while a target is
  configured (26 h daily / 170 h weekly bound)

collab:
- maintenance listener: on enter, persist + close every live session and
  refuse new connections until exit (failsafe timeout 30 min) — no
  in-memory document may write pre-restore content back afterwards

web:
- Admin → System backup section: status card with remote facts and a
  "Back up now" button, the Nextcloud settings form with test button,
  and the restore picker (local + remote sets, type-to-confirm)
- global maintenance screen: any 503 maintenance_mode flips the SPA to a
  status page polling the exempt endpoint, reloading when the instance
  returns

Verified end-to-end against a live stack (fresh DB, native api + sidecar,
fake WebDAV server): configure → test → manual backup → bundle upload →
readyz/sets/status surfaces → remote restore with maintenance gate,
marker rollback and api restart; suites: shared 21, backup 9, collab 11,
api 58 files green, lint + i18n:check + typecheck clean.

Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
Claude-Session: https://claude.ai/code/session_01EwZ4jR4KFAPvpjWevfUGX1
2026-07-12 10:39:18 +02:00

174 lines
5.2 KiB
TypeScript

import { z } from 'zod';
import type { BackupStatus, RestoreStatus } from './backup-status';
/**
* Site-Admin system panel (issue #86): maintenance jobs, backup status,
* audit trail, and storage overview — the operator's single glance for
* instance health (operations.md §Maintenance jobs).
*/
export interface SystemJobView {
name: string;
cadenceSeconds: number;
status: 'IDLE' | 'RUNNING' | 'FAILED';
lastRunAt: string | null;
lastDurationMs: number | null;
lastError: string | null;
/** False for a database row whose job no longer registers in this build. */
registered: boolean;
}
export type JobTriggerOutcome = 'succeeded' | 'failed' | 'already_running';
export interface JobTriggerResult {
outcome: JobTriggerOutcome;
job: SystemJobView;
}
export interface SystemBackupView {
/** False when no status.json exists (sidecar never ran / not deployed). */
available: boolean;
/** Freshness verdict mirroring the readyz `backup` check (issue #85). */
fresh: boolean;
status: BackupStatus | null;
maxAgeHours: number;
/** Whether a Nextcloud target is fully configured (issue #103). */
remoteConfigured: boolean;
/** Progress/result of the last in-app restore, if any (issue #103). */
restore: RestoreStatus | null;
}
/**
* Backup configuration surface of the admin settings page (issue #103).
* The app password is write-only: the view only says whether one is stored.
*/
export interface BackupSettingsView {
localRetentionDays: number | null;
remoteRetentionDays: number;
nextcloud: {
enabled: boolean;
baseUrl: string;
username: string;
folder: string;
uploadSchedule: 'off' | 'daily' | 'weekly';
passwordSet: boolean;
};
}
export const backupSettingsInputSchema = z.object({
/** `null` = no override; the sidecar's env value stays authoritative. */
localRetentionDays: z.number().int().min(1).max(3650).nullable(),
remoteRetentionDays: z.number().int().min(1).max(3650),
nextcloud: z.object({
enabled: z.boolean(),
baseUrl: z.string().trim().url({ message: 'validation.url' }).or(z.literal('')),
username: z.string().trim().max(200),
folder: z.string().trim().min(1).max(500),
uploadSchedule: z.enum(['off', 'daily', 'weekly']),
/** Empty or omitted = keep the stored app password. */
password: z.string().max(500).optional(),
}),
});
export type BackupSettingsInput = z.infer<typeof backupSettingsInputSchema>;
export const backupConnectionTestInputSchema = z.object({
baseUrl: z.string().trim().url({ message: 'validation.url' }),
username: z.string().trim().min(1).max(200),
folder: z.string().trim().min(1).max(500),
/** Empty or omitted = test with the stored app password. */
password: z.string().max(500).optional(),
});
export type BackupConnectionTestInput = z.infer<typeof backupConnectionTestInputSchema>;
export interface BackupConnectionTestResult {
ok: boolean;
/** Raw transport/server error for the admin, like the SMTP test (#80). */
error?: string;
}
/** One restorable set as offered in the restore picker (issue #103). */
export interface BackupSetView {
backupId: string;
/** UTC run start derived from the backup id. */
startedAt: string;
/** Bundle size (remote) or dump+archive sum (local); null when unknown. */
sizeBytes: number | null;
}
export interface BackupSetsView {
local: BackupSetView[];
remoteConfigured: boolean;
remote: BackupSetView[];
/** Present when a configured remote target could not be listed. */
remoteError?: string;
}
export const backupIdSchema = z.string().regex(/^\d{8}-\d{6}$/);
export const backupRestoreInputSchema = z.object({
source: z.enum(['local', 'remote']),
backupId: backupIdSchema,
/** Type-to-confirm safety: must repeat the backup id verbatim. */
confirm: z.string(),
});
export type BackupRestoreInput = z.infer<typeof backupRestoreInputSchema>;
/** Response of the public, maintenance-exempt restore status endpoint. */
export type RestoreStatusResponse = { state: 'idle' } | RestoreStatus;
export interface AuditActorView {
id: string;
username: string;
displayName: string;
}
export interface AuditEntryView {
id: string;
at: string;
action: string;
actor: AuditActorView | null;
targetType: string | null;
targetId: string | null;
details: Record<string, unknown> | null;
}
export const AUDIT_PAGE_SIZE = 50;
export const auditListQuerySchema = z.object({
/** Exact username of the acting user. */
actor: z.string().trim().min(1).optional(),
/** Action id or prefix, e.g. `grant.` matches created and deleted. */
action: z.string().trim().min(1).optional(),
from: z.coerce.date().optional(),
to: z.coerce.date().optional(),
page: z.coerce.number().int().min(1).default(1),
});
export type AuditListQuery = z.infer<typeof auditListQuerySchema>;
export interface AuditListView {
entries: AuditEntryView[];
page: number;
pageCount: number;
total: number;
}
export interface StoragePondView {
pondId: string;
name: string;
slug: string;
/** Lowercase like PondView's `type` (the api's wire convention). */
type: 'personal' | 'shared';
storageBytesUsed: number;
}
export interface StorageOverviewView {
totalBytes: number;
/** Top ponds by storage use, largest first. */
ponds: StoragePondView[];
}