Summary
腾讯云语音识别 ASR Skill,适用于语音转文字、音频转写、字幕生成、会议转录、语音消息识别、 本地文件或 URL 音频识别。包含三种模式:一句话识别(<=60s 短音频)、录音识别极速版 (<=2h/100MB 中长音频快速同步返回)、录音识别(<=5h 长音频异步识别)。支持普通话、…
tencentcloud/tencentcloud-speech-skills · Archived
npx skills add https://github.com/tencentcloud/tencentcloud-speech-skills
腾讯云语音识别 ASR Skill,适用于语音转文字、音频转写、字幕生成、会议转录、语音消息识别、 本地文件或 URL 音频识别。包含三种模式:一句话识别(<=60s 短音频)、录音识别极速版 (<=2h/100MB 中长音频快速同步返回)、录音识别(<=5h 长音频异步识别)。支持普通话、…
This repository is archived — consider an actively maintained alternative.
腾讯云语音合成(TTS)服务技能包。当用户需要将文本转换为语音文件时使用此技能,支持多种音频格式输出…
1 installsStage 1 of Clinical ASR Flywheel. Use when bootstrapping a cycle: NVCF+MW disclosure, NVIDIA_AP…
1.8K installsStage 4 of the Clinical ASR Flywheel. Use when priority KER is above 0.3 to run stock NeMo SFT …
1.8K installsStage 2 of the Clinical ASR Flywheel. Use when curating clinical terms, tagging IPA, and synthe…
1.8K installsRelated neighbors and high-traction skills in the same topics — useful to compare before installing.
Action recognition from video sequences. Supports RGB, optical flow, and joint (multi-stream) i…
1.6K installsMetric-learning recognition (ml-recog) for fine-grained visual recognition. Learns embeddings f…
1.5K installs>- Use this skill whenever the user wants text extracted from images, photos, scans, screenshot…
4K installsTranscribe speech to text using Apple's Speech framework.
3.2K installs>- Use this skill whenever the user wants text extracted from images, photos, scans, screenshot…
562 installsSpot patterns appearing in 3+ domains to find universal principles
426 installsOther skills from tencentcloud/tencentcloud-speech-skills.
npx skills add https://github.com/tencentcloud/tencentcloud-speech-skills
Declared targets from SKILL.md / docs. Unmarked agents are not listed — the skill may still install via the CLI.
main