자동 생성 — 편집 전 · 숫자만 확인된 페이지입니다
- TESIGN / RADAR
- 저장소 카드
- openbmb/voxcpm
openbmb/voxcpm
VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning
순위 역대 187위
한눈에 보는 사실
- 언어
- Python
- 라이선스
- Apache-2.0
- 사용 범위
- 상업 이용·수정·재배포 가능. 고지 유지, 바꾼 부분은 표시.
- 활동
- 마지막 커밋 14일 전 ()
- 토픽
- audio
- deeplearning
- minicpm
- multilingual
- python
- pytorch
- speech
- speech-synthesis
- text-to-speech
- tts
- tts-model
- voice-cloning
- 홈페이지
- https://voxcpm.com
- 저장소
- GitHub ↗
- 분류
- AI
요약이며 법적 조언이 아닙니다.
README 발췌
VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning English 中文 👋 Join our community for discussion and support! Feishu Discord 📚 MiniCPM Wiki VoxCPM is a tokenizer-free Text-to-Speech system that directly generates continuous speech representations via an end-to-end diffusion autoregressive architecture , bypassing discrete tokenization to achieve highly natural and expressive synthesis. VoxCPM2 is the latest major release — a 2B parameter model trained on over 2 million hours of multilingual speech data, now supporting 30 languages , Voice Design , Controllable Voice Cloning , and 48kHz studio-quality audio output. Built on a MiniCPM-4 backbone. ✨ Highlights - 🌍 30-Language Multilingual — Input text in any of the 30 supported languages and synthesize directly, no language tag needed - 🎨 Voice Design — Create a brand-new voice from a natural-language description alone (gender, age, tone, emotion, pace …), no reference audio required - 🎛️ Controllable Cloning — Clone any voice from a short reference clip, with optional style guidance to steer emotion, pace, and expression while preserving the ori…
GitHub README의 앞부분을 저장된 그대로 최대 1,200자까지 옮겼습니다. 마크다운 서식은 표시하지 않습니다.
타임라인
- 저장소 생성
- 마지막 push
- TESIGN이 처음 본 시각 ◌ 과거 기록
목록은 운영자의 위임을 받아 AI가 고릅니다. 2026년 9월에는 사람이 직접 검수하지 않습니다. 별 총합·증가량·교차 출처·마지막 업데이트·라이선스를 판단 근거로 보여줍니다. 순위 계산 방법 →
이 저장소에 편집 문구(왜·빌드·누가·시작·주의)가 붙으면 정식 항목으로 승격됩니다. 그때까지는 저장된 GitHub 메타데이터와 숫자만 보여줍니다.