tooling/db/ (a misnamed directory: connectors, not database work),
trellis-batch.sh and synth_ui_sounds.py become `reach assets`:
audio {health,generate,batch,post {convert,normalize,trim,pipeline}},
image {health,generate}, trellis {health,generate,batch}, and synth-ui.
The four audio bash wrappers are retired, and tooling/db/ is gone.
Parity, from baselines taken before anything moved:
- the four UI-sound WAVs and the harmonic-synth WAVs (exponential and linear
decay) are byte-identical
- the ffmpeg pipeline's decoded PCM is identical. Its .ogg bytes are not,
even between two runs of the OLD code: Ogg picks a random stream serial,
so the encoded file was never the right thing to compare
- the network success paths can't be run in a gate (Stable Audio and Trellis
are kept off, Gemini costs money), so tooling/test_assets.py stands up a
fake Gradio and pins every payload: the audio submit, Trellis's six-call
session sequence with its 9-input image_to_3d, and the Gemini body. It
failed when one Trellis value was mutated (7.5 → 7.0)
Failure classification, in endpoints.py, is the point of the port. The
services are OFF by design (VRAM on tower-of-joy, D-17), and the topology doc
warns against "fixing" one by restarting it. So a refused connection says OFF
and asks for the service to be turned on rather than restarted; a 4xx/5xx says
the request was rejected; 401/403 says credentials; 429 says quota; and an
unreachable Gemini blames the network, not VRAM.
Behaviour changes, each a failure that used to read as success or crash:
- audio batch and trellis batch exited 0 with failures in their summaries;
they now print the summary and exit 1
- trellis generate on a missing image crashed with a TypeError
(print(..., indent=2)); it now names the file, and checks it before the
service so a typo is not reported as an outage
- the ffmpeg pipeline left its intermediates behind when a step failed
Structure: the connectors called each other as subprocesses (batch spawned
the connector, which spawned audio_post) and parsed each other's stdout. They
are now function calls, and ffmpeg is the only exec, through core/process.
ensure_venv() is removed: it os.execv'd into .venv, which D-263's exec rule
forbids, and reach declares the dependencies itself. config.json moved into
the domain deliberately, and the local-services rule follows it.
Output contract: results are still JSON on stdout with the same keys, so skill
readers keep working. Failures are an exit status with a Fix line, never
{"ok": false}. The audio-gen, glb-gen and image-gen skills, Araminta's agent
file and the allow-list are updated to match. glb-gen's "trellis-batch.sh is
hardcoded to one category" caveat is gone: batch takes --input-dir or --names.
Co-Authored-By: Claude Opus 5.5 <noreply@anthropic.com>
title, description, type, status
| title | description | type | status |
|---|---|---|---|
| Audio Asset Pipeline | Audio pipeline index with category status, generation tools, bus architecture, and sprint scope | design | active |
Audio Asset Pipeline
Status: Active — Sprint 7 (first audio sprint)
Summary
| Category | File | Total | Planned | In-Progress | Placeholder | Final |
|---|---|---|---|---|---|---|
| Ambient | ambient.md | 4 | 4 | 0 | 0 | 0 |
| SFX | sfx.md | 2 | 2 | 0 | 0 | 0 |
| UI | ui.md | 6 | 4 | 0 | 0 | 0 |
| Total | 12 | 10 | 0 | 0 | 0 |
Update counts when asset status changes.
Palette
See palette.md for the sonic identity: two sonic families (insert-tech / organic), station palette (functional warmth), generation approach, bus architecture.
Generation Tools
- Stable Audio Open: Self-hosted, >200ms organic sounds. ~47s ceiling per generation.
- Manual synthesis: <200ms insert-tech sounds. Audacity / Python oscillator.
- Post-processing: LUFS normalization, EQ carving, crossfade loops (3-5s overlap).
Bus Architecture
5 player-facing buses (see palette.md for full spec):
| Bus | Content |
|---|---|
| Music | Future — empty |
| Ambient | Station hum, zone overlays |
| World SFX | NPC footsteps, doors, conversation murmur |
| Player Actions | Player footsteps, combat |
| UI Sounds | Cursor, implant, chimes, recognition |
Sprint Scope
Sprint 7 (#440 — 8 assets)
- 4 UI sounds: cursor hover, fog recognition, implant open, weapon aim (sketch)
- 2 UI sounds: monologue chime normal + urgent (PLACEHOLDER — Sprint 8 redo)
- 2 mix specs: dialogue dip, confrontation dip (documentation, not audio files)
Sprint 8+ (D-038 — 6 remaining assets)
- 4 ambient loops: station base, workplace layer, bar layer, corridor layer
- 2 SFX: footstep metal walk, footstep metal run
- Monologue chime redo (priority #1)
Decision References
- D-038: Audio in v0.1 scope
- D-018: Three-range sound model
- D-045: Environmental neutrality