kazeia/scripts
Kazeia Team 10a3904d7d Multi-segment TTS for long text: split → generate → concatenate
- prepare_tts_segments.py: splits text at sentence boundaries,
  generates Python pre-computed embeds per segment
- Kotlin: detects multi-segment file format, processes each segment
  independently (fresh KV cache), concatenates audio
- Long text tested: 3 segments, 335 tokens, 26.8s audio, RTF 1.67

File format: n_segments, then per segment: nPrefill, nTotal, embeds[]

Co-Authored-By: Claude Opus 4.6 (1M context) <noreply@anthropic.com>
2026-04-09 14:34:05 +02:00
..
cp_et_runner.cpp Initial commit: Kazeia TTS pipeline on NPU via ExecuTorch 2026-04-09 08:42:11 +02:00
export_cp_pte.py Initial commit: Kazeia TTS pipeline on NPU via ExecuTorch 2026-04-09 08:42:11 +02:00
export_talker_pte.py Reduce talker KV_LEN 100→64: saves 148ms (RTF 1.31) 2026-04-09 12:47:30 +02:00
prepare_tts_embeds.py Add prepare_tts_embeds.py for any text + codec_sum fix 2026-04-09 14:05:42 +02:00
prepare_tts_segments.py Multi-segment TTS for long text: split → generate → concatenate 2026-04-09 14:34:05 +02:00
qc_schema_serialize_patched.py Initial commit: Kazeia TTS pipeline on NPU via ExecuTorch 2026-04-09 08:42:11 +02:00
test_cp_et_quality.py Initial commit: Kazeia TTS pipeline on NPU via ExecuTorch 2026-04-09 08:42:11 +02:00