generated from Labyricorn/labyricorn-project-template
docs: update documentation and devlog for TalkBox fork and rebrand
CI / frontend-quality (push) Canceled after 0s
CI / frontend-quality (push) Canceled after 0s
- Updated CHANGELOG.md: added fork, rebrand, port isolation, and DCP workflow entries to Unreleased - Updated project record: set TalkBox project identity, summary, and architecture tags - Updated devlog index: set TalkBox devlog title and summary - Added devlog entry 'fork-and-rebrand': initial fork, rebrand, port migration, and devlog workflow
This commit is contained in:
@@ -2,26 +2,34 @@ _model: project
|
||||
---
|
||||
schema_version: 1
|
||||
---
|
||||
project_id: labyricorn-project-template
|
||||
project_id: talkbox
|
||||
---
|
||||
title: Labyricorn Project Template
|
||||
title: TalkBox
|
||||
---
|
||||
summary:
|
||||
|
||||
Initial template and publishing structure for Labyricorn-compatible projects and development logs.
|
||||
Local-first AI voice studio for multi-engine speech generation, zero-shot voice cloning, global dictation, and MCP agent voice I/O.
|
||||
---
|
||||
status: active
|
||||
---
|
||||
started: 2026-08-23
|
||||
started: 2026-08-24
|
||||
---
|
||||
author: Labyricorn
|
||||
---
|
||||
repository_url: https://git.labyricorn.com/Labyricorn/labyricorn-project-template
|
||||
repository_url: https://git.labyricorn.com/Labyricorn/TalkBox
|
||||
---
|
||||
default_branch: main
|
||||
---
|
||||
tags: template, lektor, python
|
||||
tags: tts, voice-cloning, stt, whisper, mcp, tauri, react, fastapi, rust, python
|
||||
---
|
||||
body:
|
||||
|
||||
Starter template providing the standard .labyricorn/ publishing records, assistant instructions, devlog structure, and repository tooling for publication on Labyricorn.
|
||||
TalkBox is a local-first AI voice studio combining high-fidelity speech synthesis, zero-shot voice cloning, global dictation with synthetic paste, and Model Context Protocol (MCP) voice I/O.
|
||||
|
||||
### Highlights
|
||||
|
||||
- **Multi-Engine Speech Generation**: Supports 7 TTS engines (Qwen3-TTS, Qwen CustomVoice, LuxTTS, Chatterbox Multilingual, Chatterbox Turbo, HumeAI TADA, Kokoro) covering 23 languages.
|
||||
- **Voice Cloning & Profiles**: Zero-shot voice cloning from reference audio samples, custom effects chains, and personality prompt attachments.
|
||||
- **Global Dictation**: Whisper-backed push-to-talk and toggle dictation anywhere on macOS, Windows, and Linux with focus-aware automatic pasting.
|
||||
- **Agent Voice Output (MCP)**: Streamable HTTP and stdio MCP server (`http://127.0.0.1:17494/mcp`) exposing `talkbox.speak`, `talkbox.transcribe`, `talkbox.list_captures`, and `talkbox.list_profiles` for agent integrations.
|
||||
- **Native & Local Performance**: Built on Tauri v2 (Rust) and React/TypeScript frontend with Python FastAPI backend running locally on Apple Silicon (MLX), NVIDIA (CUDA), AMD (ROCm), DirectML, or CPU.
|
||||
Reference in New Issue
Block a user