feat: add FunASR dataset annotation tool - #1309
Open
LauraGPT wants to merge 1 commit into
Open
Conversation
Author
|
@leng-yue, could you review this optional dataset-annotation tool when convenient? You recently merged changes in the same installation/finetuning documentation area. The current head is CLEAN and mergeable, with no unresolved review threads or failed/pending checks. FunASR stays lazily imported outside Fish Speech's core dependencies; the CLI preserves existing .lab files by default and has dry-run, overwrite, atomic-write, and per-file failure coverage. Exact-head validation includes 9 unit tests, repository pre-commit hooks, dependency-free help/dry-run smokes, and real H100 SenseVoiceSmall transcription for the official Chinese and English samples. |
This file contains hidden or bidirectional Unicode text that may be interpreted or compiled differently than what appears below. To review, open the file in an editor that reveals hidden Unicode characters.
Learn more about bidirectional Unicode characters
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment
Add this suggestion to a batch that can be applied as a single commit.This suggestion is invalid because no changes were made to the code.Suggestions cannot be applied while the pull request is closed.Suggestions cannot be applied while viewing a subset of changes.Only one suggestion per line can be applied in a batch.Add this suggestion to a batch that can be applied as a single commit.Applying suggestions on deleted lines is not supported.You must change the existing code in this line in order to create a valid suggestion.Outdated suggestions cannot be applied.This suggestion has been applied or marked resolved.Suggestions cannot be applied from pending reviews.Suggestions cannot be applied on multi-line comments.Suggestions cannot be applied while the pull request is queued to merge.Suggestion cannot be applied right now. Please check back later.
Is this PR adding new feature or fix a BUG?
Add feature.
Is this pull request related to any issue? If yes, please link the issue.
Closes #1291.
Summary
The default SenseVoiceSmall path writes plain, one-line UTF-8 transcripts for Mandarin, Cantonese, English, Japanese, and Korean. Model selection, device, language, and ITN remain configurable.
Validation
No Fish Speech runtime dependency or training behavior changes are included.