skillZs
★ LIVE SKILL TAGS ★
>>> LIVE SKILLS INDEX <<<
* OPEN SOURCE *
NO LOGIN, NO TRACKING
※ REAL INSTALL DATA ※
← back to all skills
decolua/9router108 installs

9router-stt

Speech-to-text via 9Router /v1/audio/transcriptions using OpenAI Whisper / Groq / Gemini / Deepgram / AssemblyAI / NVIDIA / HuggingFace models. Use when the user wants to transcribe audio, convert speech to text, or get subtitles from audio files.

How do I install this agent skill?

npx skills add https://github.com/decolua/9router --skill 9router-stt
view source ↗

Is this agent skill safe to install?

  • Gen Agent Trust Hubpass

    The skill provides speech-to-text functionality using the 9Router API, supporting multiple AI providers. It includes examples for model discovery and audio transcription, while referencing documentation on the developer's official GitHub repository.

  • Socketwarn

    1 alert: gptAnomaly

  • Snykpass

    Risk: LOW · No issues

What does this agent skill do?

9Router — Speech-to-Text

Requires NINEROUTER_URL (and NINEROUTER_KEY if auth enabled). See https://raw.githubusercontent.com/decolua/9router/refs/heads/master/skills/9router/SKILL.md for setup.

Discover

curl $NINEROUTER_URL/v1/models/stt | jq '.data[].id'
# Per-model params (language, response_format, prompt, temperature support)
curl "$NINEROUTER_URL/v1/models/info?id=openai/whisper-1"

model = STT model ID (e.g. openai/whisper-1, groq/whisper-large-v3, deepgram/nova-3, gemini/gemini-2.5-flash).

Endpoint

POST $NINEROUTER_URL/v1/audio/transcriptions (OpenAI Whisper compatible, multipart/form-data)

FieldRequiredNotes
modelyesfrom /v1/models/stt
fileyesaudio file (mp3, wav, m4a, webm, ogg, flac)
languagenoISO-639-1 (e.g. en, vi)
promptnohint text to guide transcription
response_formatnojson (default) / text / verbose_json / srt / vtt
temperatureno0–1

Examples

curl -X POST "$NINEROUTER_URL/v1/audio/transcriptions" \
  -H "Authorization: Bearer $NINEROUTER_KEY" \
  -F "model=openai/whisper-1" \
  -F "file=@audio.mp3" \
  -F "language=vi"

JS (Node):

import { createReadStream } from "node:fs";
const form = new FormData();
form.append("model", "groq/whisper-large-v3-turbo");
form.append("file", new Blob([await (await import("node:fs/promises")).readFile("audio.mp3")]), "audio.mp3");
const r = await fetch(`${process.env.NINEROUTER_URL}/v1/audio/transcriptions`, {
  method: "POST",
  headers: { "Authorization": `Bearer ${process.env.NINEROUTER_KEY}` },
  body: form,
});
const { text } = await r.json();
console.log(text);

Response shape

Default (response_format=json):

{ "text": "Xin chào, đây là bản ghi âm." }

verbose_json adds language, duration, segments[] with timestamps. srt / vtt return subtitle text.

Provider quirks

Providermodel formatNotes
openaiwhisper-1, gpt-4o-transcribe, gpt-4o-mini-transcribeNative OpenAI shape
groqwhisper-large-v3, whisper-large-v3-turbo, distil-whisper-large-v3-enFastest; OpenAI shape
geminigemini-2.5-flash, gemini-2.5-pro, gemini-2.5-flash-liteServer converts to generateContent with audio inline
deepgramnova-3, nova-2, whisper-largeToken auth; server adapts response
assemblyaiuniversal-3-pro, universal-2Async upload+poll handled server-side
nvidianvidia/parakeet-ctc-1.1b-asrNIM endpoint
huggingfaceopenai/whisper-large-v3, openai/whisper-smallHF Inference API

Add the canonical catalog link to the repository README so users can inspect current installs and available audits. The publishing guide covers the complete discovery path.

<a href="https://skillzs.dev/skills/decolua/9router/9router-stt">View 9router-stt on skillZs</a>