Update README.md
Browse files
README.md
CHANGED
|
@@ -12,6 +12,11 @@ tags:
|
|
| 12 |
|
| 13 |
TADA is a unified speech-language model that synchronizes speech and text into a single, cohesive stream via 1:1 alignment. By leveraging a novel tokenizer and architectural design, TADA achieves high-fidelity synthesis and generation with a fraction of the computational overhead required by traditional models.
|
| 14 |
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
| 15 |
## Key Features
|
| 16 |
|
| 17 |
- 1:1 Token Alignment: Unlike standard models, TADA’s tokenizer encodes audio into a sequence of vectors that perfectly matches the number of text tokens.
|
|
@@ -58,6 +63,9 @@ We provide several model checkpoints:
|
|
| 58 |
|
| 59 |
All models use the same encoder ([`HumeAI/tada-codec`](https://huggingface.co/HumeAI/tada-codec)) and can be loaded using the same API.
|
| 60 |
|
|
|
|
|
|
|
|
|
|
| 61 |
## Run Inferece
|
| 62 |
|
| 63 |
### Text-to-Speech
|
|
|
|
| 12 |
|
| 13 |
TADA is a unified speech-language model that synchronizes speech and text into a single, cohesive stream via 1:1 alignment. By leveraging a novel tokenizer and architectural design, TADA achieves high-fidelity synthesis and generation with a fraction of the computational overhead required by traditional models.
|
| 14 |
|
| 15 |
+
⭐️ arxiv: https://arxiv.org/abs/2602.23068
|
| 16 |
+
⭐️ demo: https://huggingface.co/spaces/HumeAI/tada
|
| 17 |
+
⭐️ github: https://github.com/HumeAI/tada
|
| 18 |
+
⭐️ blog post:
|
| 19 |
+
|
| 20 |
## Key Features
|
| 21 |
|
| 22 |
- 1:1 Token Alignment: Unlike standard models, TADA’s tokenizer encodes audio into a sequence of vectors that perfectly matches the number of text tokens.
|
|
|
|
| 63 |
|
| 64 |
All models use the same encoder ([`HumeAI/tada-codec`](https://huggingface.co/HumeAI/tada-codec)) and can be loaded using the same API.
|
| 65 |
|
| 66 |
+
## Evaluation
|
| 67 |
+
|
| 68 |
+
|
| 69 |
## Run Inferece
|
| 70 |
|
| 71 |
### Text-to-Speech
|