ClearerVoice-Studio
SkillProductivity"Route ClearerVoice-Studio tasks for ClearVoice inference,
Available today. Use it from your connected AI after setup.
No other account needed.
Connect ahel once, and every AI you use reads what you have installed.
Then ask your AI: use the ClearerVoice-Studio skill
What this skill tells your AI
The instructions your AI receives, as published by vectorspacelab/arex-skill in skills/repositories/repo-skills/clearer-voice-studio/SKILL.md and read by ahel’s review.
Use this repo skill when a task involves ClearerVoice-Studio, the clearvoice package, ClearVoice pretrained speech processing models, SpeechScore quality metrics, or the repository's training and data-preparation workflows.
Fast route map
| User task | Read |
|---|---|
| Use pretrained models for speech enhancement, speech separation, speech super-resolution, or audio-visual target speaker extraction | sub-skills/clearvoice-inference/SKILL.md |
Choose task/model_names, validate file/directory/.scp inputs, handle NumPy/Tensor inference, or debug checkpoint/media issues | sub-skills/clearvoice-inference/SKILL.md |
| Compute PESQ, STOI, SI-SDR, SNR, DNSMOS, NISQA, DISTILL_MOS, or other SpeechScore metrics | sub-skills/speechscore-metrics/SKILL.md |
| Score matched directories, decide whether a clean reference is required, or troubleshoot metric dependencies | sub-skills/speechscore-metrics/SKILL.md |
Train, fine-tune, resume, run training-side inference, prepare .scp/CSV data lists, or generate noisy/reverb training data | sub-skills/training-and-data-prep/SKILL.md |
| Check whether a local environment is ready without downloading weights or launching training | scripts/check_clearer_voice_environment.py |
Component overview
Read references/package-and-component-overview.md when you need the repository component map, verified public signatures, package/source layout notes, or high-level prerequisites.
Key facts:
clearvoiceis the installable package for ClearVoice inference:from clearvoice import ClearVoice.- The public ClearVoice constructor is
ClearVoice(task, model_names). - Calling a
ClearVoiceinstance with a string path uses file/directory/.scpI/O mode; calling with a NumPy array or Torch tensor uses tensor-to-tensor mode. - SpeechScore is a source-layout component in this snapshot; use its helper or source-layout notes when importing it outside its component directory.
- Training workflows are script/config based and should be treated as templates that require user-owned datasets, checkpoints, GPUs, and explicit output paths.
Safe first checks
Before running expensive or mutating workflows:
- Run
python scripts/check_clearer_voice_environment.pyto checkclearvoice, Torch/CUDA visibility, and FFmpeg presence without loading model weights. - For SpeechScore source-layout checks, pass the user's SpeechScore component directory to
--speechscore-dir. - Use each sub-skill's dry-run helper before real model inference, metric scoring, or training launch preparation.
- Ask before starting model downloads, distributed training/evaluation, data generation, checkpoint overwrites, or package backend changes in a user-owned environment.
Installation guidance
For pretrained ClearVoice package use:
pip install clearvoice
python - <<'PY'
from clearvoice import ClearVoice
print(ClearVoice)
PY
For full repository workflows such as SpeechScore metrics and training scripts, install the repository's documented runtime requirements in an isolated environment. Do not modify a user's existing working environment without approval.
Troubleshooting
- Read references/troubleshooting.md for cross-cutting install/import, CUDA, FFmpeg, model-download, and component-routing failures.
- Use sub-skill troubleshooting references for workflow-specific symptoms and recovery steps.
- Read references/repo-provenance.md before deciding whether this skill is stale for a different checkout or package version.
Boundaries
Use this skill to operate the package/repository workflows, not to modify the ClearerVoice-Studio source code. If the user asks for generic speech recognition, TTS, diarization, or unrelated PyTorch training, route to a more specific skill. If the user asks to edit this repository's source, treat it as a repository-maintenance task and inspect the current checkout directly.
Signals
- GitHub stars
- 266
- Forks
- 21
- Last commit
- Sep 2026
ahel review
K1binfo
installs-packagesK6low
bundled executables the agent is told to runK1binfo
installs-packages (in sub-skills/clearvoice-inference/scripts/clearvoice_inference_recipe.py)K1binfo
installs-packages (in sub-skills/clearvoice-inference/scripts/clearvoice_numpy_recipe.py)K1binfo
installs-packages (in references/troubleshooting.md)K1binfo
installs-packages (in sub-skills/clearvoice-inference/references/troubleshooting.md)K1binfo
installs-packages (in sub-skills/speechscore-metrics/references/troubleshooting.md)
Automated review, not a security audit. Ruleset v1+k2.
Advanced
- Catalog kind
- skill
- Gateway key
clearer-voice-studio- Source
- github.com/vectorspacelab/arex-skill