The mirror's deploy pair had drifted ~280 lines behind: it lacked the
transactional .env preserve/restore (an interrupted deploy could destroy
server-side env files), the stderr-flattening step wrapper (a successful
deploy reported failure and skipped its own verification), and the
verification rework (channel re-picked every retry, edge only with a Host
to route by, verify.timeoutSeconds, honest split of "stale build" vs "no
channel answered").
The sync is byte-faithful to the private tree except where the mirror's
own Sanitization suite demands otherwise - and it caught the first copy:
three failures for private project names, a private domain, and a
private-only script name that rode along in comments. Each war story keeps
its lesson and loses its cast, per the convention already in the file
("EvoCivilCode: deploy/" was already published as "(deploy/, infra/,
...)"). That suite going red on an unsanitized copy is exactly what it
exists for.
Also in this change:
- tests/VerifyPlan.Tests.ps1 - the channel-selection rules are pure
functions and Pester pins them (no domain => no edge attempt; the PS 5.1
one-element-unroll trap). First verification tests in the mirror.
- README: the verify block now documents viaProxy/upstream (they shipped
in the docker-network read but were never in the README),
timeoutSeconds, and the channel order with why it re-resolves per retry.
- README.txt: generated plain-text twin, via scripts/readme_txt.py
(vendored from the fleet's reference implementation; the file is
generated, never edited by hand).
- CHANGELOG.md/.txt: entries merged into the existing Unreleased sections.
CHECKSUMS.txt refreshed (42 entries). Pester: 240 passed, 0 failed.
Both synced files parse clean.
Observed, untouched: the Unreleased section carries duplicate "### Fixed"
headings from earlier appends; folding them risks reordering entries whose
prose references their neighbours, so it is left for the next release cut
(zbump #110 rolls Unreleased into the version being cut).
13 KiB
Changelog
Notable changes to the Evomedia.net Token Savers.
Unreleased
Changed
zdeployverification picks its channel on every retry, and never asks the bare IP — the check used to choose its channel once, before the wait loop, by probing; the probes raced the app restart the check exists to wait through, so the whole window went to the edge fallback. For a project with nodomainthat fallback had noHostheader, and the proxy can then only answer from its default vhost — a different product: that is how one deploy's check compared another app's build number against its own. Now channels are re-resolved each retry in trust order (docker-networkviaProxy→localhost:<port>→ edge with the project'sHost), the edge is skipped entirely when there is no host to route by, and a project with no trustworthy channel is reported as unverifiable instead of guessed at. The final warning also says which failure happened: a version that never matched (stale/failed build) reads differently from channels that never answered (probably still booting).verify.timeoutSecondsjoins the config so a project that is slow to boot — e.g. one that runs database migrations in its entrypoint — can widen its own window instead of warning on every routine success.- An interrupted deploy can no longer destroy server-side
.envfiles — the preserve/restore of operator files is transactional: the restore comes from a tarball taken before the tree is replaced, so a deploy that dies mid-flight leaves the previous files in place instead of an empty directory. - A successful deploy no longer reports failure —
docker compose restartwrites routine progress to stderr, which PowerShell 5.1 turns into a terminating error under$ErrorActionPreference = 'Stop'; four ssh calls bypassed the wrapper that flattens this. All remote steps now run through it and are judged by exit code alone. zversion bumpis once per release, not once per PR — the usage text and thebumphelp line both said "one per PR, one per defect fix". The build number names something that shipped, so a release carrying five PRs moves it by one; PRs that never shipped on their own were never separate builds. Help text only here, but it is the wording people follow: it stamped a single evo.www release as two builds. The historical entry below, which records what the rule was whenzversionshipped, is deliberately left as written.
Added
-
tests/VerifyPlan.Tests.ps1— the verification channel-selection rules are pure functions inZHelpers.ps1(Get-VerifyAttempts,Get-VerifyTimeout) and Pester pins them, including "a project with no domain must never produce an edge attempt" and the PowerShell 5.1 one-element-unroll trap. -
scripts/readme_txt.pyandREADME.txt— a generated plain-text twin of the README for terminals and pagers.README.txtis generated, never edited by hand. -
Read a live build from inside the docker network, not through the public proxy —
zdeploy,zec2andzec2onlinenow preferdocker exec <viaProxy> curl http://<upstream>/api/build-versionwhen a project setsverify.viaProxyandverify.upstream.Two problems it closes. A build stamp is something many sites deliberately do not serve publicly, and a checker that reads it over the public URL stops working the moment that endpoint is blocked — reporting "unknown", which is indistinguishable from "could not reach it". And the proxy answers from whichever vhost matches the Host header, so a container with no public route was getting another site's version back and failing deploys that had worked.
Reading it from a container on the shared network also exercises the real HTTP path, so it proves the app is serving rather than that its database knows a version. Purely additive: projects without those two keys behave exactly as before.
Fixed
scpno longer receives an ssh-only flag — the stdin-hang fix added-ntoGet-Ec2SshOpts, and the deploy path splats that same array intoscpas well asssh. OpenSSH'sscphas no-n: it exits 1 withunknown option -- nand prints its usage block, so every upload failed.Get-Ec2ScpOptsnow supplies the shared connection options with-nfiltered out, derived fromGet-Ec2SshOptsrather than duplicated so the timeouts cannot drift apart between the two transports. All fourscpcall sites use it, including the recursive directory upload. The upload failure message also asserted "Likely server disk space" without checking; it now points atscp's own output, where the real diagnosis already was.
Fixed
- Deploys can no longer hang forever on an ssh prompt — every
deploy-path
ssh/scpnow carriesBatchMode=yesplus connect and keepalive timeouts (Get-Ec2SshOptsinZHelpers.ps1). WithoutBatchMode, ssh prompts for a passphrase or password and waits indefinitely; because the deploy pipes stderr through the pipeline, the prompt never reaches the screen and the run just stops under whatever step label printed last, with no explanation. Now it fails immediately — there is no prompt on this path worth answering.ServerAlive*bounds a session that dies mid-command (dropped VPN, sleeping laptop, rebooting host) to about a minute instead of hanging. - Vendored archives survive the archive filter — files under a
vendor/directory are exempt from the "no archives in the zip" rule. A project that vendors a dependency asvendor/*.tgzneeds it in the deploy zip; dropping it makes a Dockerfile'sCOPY vendor ./vendorfail at image build, a confusing way to learn the filter ate a build input. unzipinstall is idempotent — the remote step ranapt-get update && apt-get install -y unzipon every deploy; it now checkscommand -v unzipfirst and skips the apt round-trip when the binary is already there.zdeployedge kind now ships asset subdirectories (#42) — the edge deploy uploaded top-level files only, so a project self-hosting assets in folders (fonts/,vendor/) lost them on every deploy: docker created empty root-owned mount points and nginx served 404s from them, which shows up as fonts silently falling back and vendored JS never loading. Every subdirectory exceptnginx-logs/and.git/now ships recursively, and theensure edge dirchown is recursive so scp into docker-created root-owned dirs cannot fail.zdeployno longer deletes operator-managed files on deploy (#2) — the project-directory replacement preserved only./.env, silently destroying every other server-side file (.env.db, staged signing keys, certs) on every deploy. All.env*files at the project root are now preserved by default, plus anything listed in the newdeploy.preservearray (files or directories); the vite kind, which previously preserved nothing, gets the same protection. Found the hard way: a first production deploy of an auth service wiped its staged DB credentials and RSA signing keys.
Added
- Versioned releases:
zversion,zrelease,releases/— the toolkit now carries one version in a 5-segment scheme,v{major}.{rc}.{beta}.{alpha}.{build}.zversion bump(one per PR / defect fix) andzversion bump-stage release|rc|beta|alpha(zeroes every lower segment) rewritebuild-version.json, stamp# Version:into every script header — so a lone copied script still says which release it came from — and regenerateCHECKSUMS.txtin the same step.zreleasepackages the current version asreleases/zscripts-<version>.zipwith a.sha256beside it: one hash verifies the download, the bundledCHECKSUMS.txtverifies the extracted contents, so nobody needs to clone the repo to get a verifiable copy. Released zips are immutable —zreleaserefuses to overwrite one. zchecksums+CHECKSUMS.txt— a SHA-256 manifest covering every.ps1and.cmd, so a download can be verified before anything is run.zchecksumschecks them;zchecksums -Updateregenerates after an intentional edit. The manifest issha256sumformat, sosha256sum -c CHECKSUMS.txtworks on Linux/macOS/WSL too, and the hashes match on every platform because.gitattributespins these files to CRLF everywhere. Flags changed files, missing files, and scripts present on disk but absent from the manifest. It's an integrity check, not a signature — the manifest sits in the same repo as the code, so it catches corruption and accidental drift, not a compromised repo. A Pester test fails if the manifest ever goes stale.- Test suite (Pester) — the toolkit now has automated coverage of its own
pure logic:
Get-ArchiveExcludes(including the deploy-vs-backup rule that keeps.env/uploadsout of deploys but in backups), config and project lookups,remote.composeDirfallback, EC2 target composition, and build-label formatting. Run withInvoke-Pester .\tests(Pester 5+). Verified by mutation testing — reintroducing each historical bug turns the suite red. ZCONFIGenvironment variable — overrides the path tozconfig.json, so a run can target an alternate config. Also gives the test suite a seam for injecting a fixture.zec2_rotatekeys— safely rotate/reset server-side secrets — a new tool for when a secret leaks or a deploy overwrites a production.envwith dev values.-Rotate KEYregenerates a key on the server (openssl rand -hex 32) so the new value never leaves the box;-Set KEYtakes an operator-known value (e.g.DATABASE_URL,ADMIN_EMAIL) from a masked prompt and streams it over SSH stdin — never a command argument, never echoed. Backs the server.envup to a timestamped.bakfirst, updates the key atomically (matches or appends), auto-detectsbackend/.envfromdeploy.preserve, and with-Restartrecreates the container (up -d --force-recreate, so the new values actually load — a plain restart keeps the old environment).-WhatIfpreviews the plan without touching anything.zkill all—zkillnow acceptsall, stopping the dev server of every project that has aports.dev(edge/docker stacks with no local dev server are skipped). Brings it in line withzdeploy all/zbackup all; the one-shot "stop everything I've got running locally".zdeployserver-side health verification (verifyblock) — projects not published through the edge proxy can declare"verify": { "port": ..., "path": "/health", "expect": "..." }and the deploy is checked from the server itself (curl localhost:<port><path>over SSH) instead of hitting the public IP. Fixes a false PASS where the proxy's default vhost answered for apps that never started; projects with neitherdomainnorverifyare now reported as NOT verified.zdeployoptionaldeploy.gitPull—git pull --ff-onlyin the project root before zipping.zdeployzips the working tree and doesn't otherwise pull, so a checkout left behindoriginafter a merged PR would deploy stale code while still bumping the build number — success that changes nothing. A failed pull aborts the deploy instead.- Per-project
startconfig block —zstarthonors optional pre-start steps fromzconfig.json:"gitPull": truerunsgit pull --ff-onlyin the project root before starting (never boot a stale checkout), and"env": { ... }sets environment variables for the dev-server process. Example added tozconfig.example.json. - Switch-style argument tolerance — a leading dash on a project key is
ignored everywhere (
zdeploy -myapp==zdeploy myapp), for hands that grew up on per-project switches.
Changed
zbackup/zbackup_and_syncrequire an explicit target — running them bare now shows usage instead of quietly backing up every project;alldoes what bare invocation used to (matchingzdeploy). The scheduled task created bysetup_backup_schedule.ps1passesall— re-run it if your task was registered before this change.zbackupparses moreDATABASE_URLstyles — double/single-quoted values (Prisma convention),postgres://andpostgresql+driver://schemes, and URLs without an explicit port (defaults to 5432) all work; previously these skipped the Postgres dump with "Could not parse DATABASE_URL".zkill/ port cleanup kills the whole process tree — listeners on a project's port are now terminated children-first. Auto-reloading servers (uvicorn/watchfiles, nodemon) spawn workers that inherit the listening socket; killing only the parent left orphans serving stale code.zbackupfindsDATABASE_URLinbackend\.envtoo — projects with a frontend/backend split get their Postgres dump bundled without needing a root-level.env.
1.0.0
Initial public release: zstart / zkill / zrestart (local dev servers),
zdeploy (zip → upload → compose build → live build-version verification,
with handlers for python / vite / nextjs / edge / docker project kinds),
zec2 / zec2online / zrepair (health checks and recovery), zbackup /
zbackup_ec2 / zsync (local, server-side, and offsite backups), all driven
by a single gitignored zconfig.json.