Skip to content

feat(ui): design voices in the voice library - #12307

Draft
localai-org-maint-bot wants to merge 3 commits into
masterfrom
feat/voice-library-design-12219
Draft

localai-org-maint-bot wants to merge 3 commits into
masterfrom
feat/voice-library-design-12219

Conversation

@localai-org-maint-bot

Copy link
Copy Markdown
Collaborator

Description

Closes #12219.

Add a Design a voice option to the Voice Library creation form. Choose a speech model, enter voice instructions and sample text, then preview and save the generated reference as a reusable profile. This removes the manual download and re-upload step.

The form normalizes generated audio using the existing upload path and fills the transcript from the generation request. Editing the design requires a new sample before saving. The existing consent and reference-duration checks still apply.

Notes for Reviewers

The model selector lists installed TTS models. Users must select a model that supports voice design; models without instruction support may ignore the description. No server API changes are needed.

Validation from core/http/react-ui:

  • node node_modules/vite/bin/vite.js build passes.
  • node scripts/inline-style-gate.mjs passes at the existing baseline (512).
  • node --check e2e/voice-library.spec.js and git diff --check pass.
  • A temporary Node harness executes the actual JSX component handlers with mocked React hooks, speech responses, and audio decoding. It failed before the change and passes afterward, checking request data, preview state, stale samples, retries, and saved multipart fields. This does not replace browser testing.
  • Five Playwright cases cover generation/save, edits and in-flight state, retry, missing models, and duration validation. Existing upload coverage remains.

Browser checks remain outstanding. Cached Chromium cannot launch in this Alpine runner (ENOENT, missing glibc), and the network proxy rejects package downloads. CI or a browser-capable host should run make test-ui-e2e before this draft becomes ready.

Signed commits

  • Yes, I signed my commits.
  • Documentation updated (docs/content/) for user-facing changes, or not applicable

Generate a reference from voice instructions and sample text, then
preview and save it as a reusable profile. Clear stale samples when
inputs change.

Assisted-by: Codex:GPT-6
@localai-org-maint-bot
localai-org-maint-bot force-pushed the feat/voice-library-design-12219 branch from 8961599 to 71f653e Compare October 8, 2026 22:04
Let model descriptions wrap without losing the install action offscreen.
Give page table rules precedence over the shared table styles.

Separate background facet probes from navigation assertions and wait for
asynchronous size sorting and masonry layout. Fix the chat fixture clock
so runs near midnight keep the expected conversation groups and order.

Assisted-by: Codex:gpt-6
Playwright joins adjacent table-cell text. A llama-cpp backend followed
by 99 requests then matches p99, although no percentile is shown.
Require word boundaries while keeping the absence assertion.

Assisted-by: Codex:gpt-6

This branch has not been deployed

No deployments
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

Voice Design Model Support in Voice/Personality Library

1 participant