- App Routing's output-voice field gets the same searchable
avatar-thumbnail dropdown used elsewhere, as a browse button
alongside the existing free-text input (which must stay editable to
target vd_ Voice Design presets not in the voice library).
- My Voices table: Gender and Rating filter dropdowns existed in the
JS (populateLibraryFilters, libraryFilterMatch) but their <select>
elements had been dropped from the visible layout after an earlier
redesign, replaced with hidden dead placeholders just to keep the
code from erroring - and since populateLibraryFilters() early-returns
if any of the three elements are missing, this silently broke the
already-visible Language/Type dropdowns too. Restored the real
elements and removed the hidden scaffold; added new Tag and Group
dropdowns wired to the same filter state the sidebar chips use.
- Audio effects failing with "pedalboard is not installed" despite
requirements.txt listing it: the package WAS installed, but its
native extension (pedalboard_native) links against libatomic.so.1,
an OS-level shared library missing from the python:3.11-slim-bookworm
base image. Added libatomic1 to the Dockerfile and rebuilt - verified
`import pedalboard` now succeeds in the running container.
- Relabeled "edit ID"/"copy ID" to "rename filename"/"copy filename"
in the voice inspector - the feature already renamed the underlying
.wav/.meta.json/.reference.txt/picture files via the existing
/api/voice/rename endpoint, it just wasn't obvious "ID" meant
"filename."
Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
- Add a 6-stage pipeline stepper (Source -> Cast Audiobook -> Cast Characters
-> Script Rehearser -> Generate MP3s -> Audiobook) with direct, non-destructive
jumps between stages and a prominent guided-tour look
- Split PDF import into an explicit "load" then "Extract Text" step, with
in-browser OCR (Tesseract.js, vendored) to recover chapter headlines baked
into a PDF as images instead of real text
- Fix casting feed silently merging pages after leaving/returning: segments
now carry their own page number instead of re-guessing it from text
- Fix excessive "Unknown" speaker attribution: restore the attribution LLM's
output token budget, which had been cut roughly in half and was truncating
dialogue-dense passages
- Fix Theater Play library cards failing to open (dead pre-migration
IndexedDB API calls, missing section navigation)
- Fix bulk "Set tag" wiping a voice's existing tags instead of adding to them
- Start merging Casting's feed with Script Rehearser's Stage UI: collapsible
character sidebar, shared "paper" page styling, inline text editing
- Fix a performance regression from that merge (per-row listeners on every
redraw) by moving to event delegation
- Various layout/clutter fixes: hide reader chrome until a document is
loaded, collapse secondary settings by default, fix overlapping toolbar
icons, fix duplicate "opening" notifications
Co-Authored-By: Claude Sonnet 5 <noreply@anthropic.com>
Without COPY VERSION, _read_version() fell back to "0.0.0". The volume
mount is a runtime override; the baked-in copy is the reliable baseline.
Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com>
- docker-compose.yml: port 7890:7890, image/container/volume all renamed to tts-voice-creator-clone-and-design-2
- Dockerfile: EXPOSE 7890
- server.py: uvicorn binds to port 7890
Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com>
- New light UI: fixed 220px sidebar, single scrolling page, 8 named sections
- Static files split by concern: style.css, app.js, loader.js, nav.js
- Each page section is its own partial in static/sections/s-*.html
- loader.js fetches all section partials in parallel, then loads app.js and nav.js
- All original functionality, element IDs, and API endpoints preserved
Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com>