$npx skillfedfor your agent

qianwen-audio-tts

Qwen Audio TTS turns written text into natural-sounding speech using Qwen's TTS engine. Choose from multiple voices and models—including the fast qwen3-tts-flash for standard tasks or instruction-guided variants for tone control—then output audio directly to file.

Qwen Audio TTS converts text to speech using Qwen's TTS models with multiple voice options and output formats.

AI-generated summary based on this skill's SKILL.md

★ 38  0 Apache-2.0updated by QianWen-AI

Decision gist · record as of 2026-06-17

Qwen Audio TTS converts text to speech using Qwen's TTS models with multiple voice options and output formats. Qwen Audio TTS turns written text into natural-sounding speech using Qwen's TTS engine. Choose from multiple voices and models—including the fast qwen3-tts-flash for standard tasks or instruction-guided variants for tone control—then output audio directly to file.

manual: git clone https://github.com/QianWen-AI/qianwen-ai → cp -r qianwen-ai/skills/audio/qianwen-audio-tts ~/.claude/skills/qianwen-audio-tts
skills/audio/qianwen-audio-tts/SKILL.md · version 8380d592

Use it when

  • Yes.
  • Qianwen-audio-tts generates voiceovers and audio narration by processing your text through Qwen's TTS engine.

Verify before relying

Read SKILL.md below before installing (11 files). Open directory: indexed for reading, not audited.

Same gist for agents: .md · .json

Install

QianWen-AI/qianwen-ai/qianwen-audio-tts · repository language: Python

Open directory. Skills are indexed for reading, not audited. Review a skill's body before installing it.

Frequently asked questions

AI-generated answers based on this skill's SKILL.md and metadata

What does qianwen-audio-tts do?

Qianwen-audio-tts converts written text into natural-sounding speech using Qwen's TTS engine. The skill lets you choose from multiple voices and models, including the fast qwen3-tts-flash for standard tasks or instruction-guided variants for tone control, then output audio directly to file.

Can I convert text to speech using qianwen-audio-tts?

Yes. Qianwen-audio-tts is built specifically to convert text to speech. It uses Qwen TTS models to synthesize natural-sounding audio from your written content, with options to select different voices and output formats for your needs.

How do I generate voiceover from text with this skill?

Qianwen-audio-tts generates voiceovers and audio narration by processing your text through Qwen's TTS engine. You can select your preferred voice and model variant, configure any tone or emotion settings if using instruction-guided models, and the skill outputs the resulting audio directly to a file.

Does qianwen-audio-tts support multiple languages?

Qianwen-audio-tts supports creating multilingual audio output from text content. This allows you to generate speech synthesis across different languages using Qwen's TTS models, making it suitable for international voiceover and narration projects.

Can I control tone and emotion in qianwen-audio-tts?

Qianwen-audio-tts includes instruction-guided model variants that enable tone and emotion control in your synthesized speech. These advanced models let you shape how the generated audio sounds beyond standard voice selection, giving you finer control over the final voiceover quality.

What license does qianwen-audio-tts use?

Qianwen-audio-tts is released under the Apache-2.0 license, which permits free use, modification, and distribution of the skill under the terms of that open-source license.

SKILL.md

Rendered from the published skill. Quoted content, verbatim.

> Agent setup: If your agent doesn't auto-load skills (e.g. Claude Code), > see agent-compatibility.md once per session.

Qwen Audio TTS (Text-to-Speech)

Synthesize natural speech from text using Qwen TTS models. This skill is part of QianWen-AI/qianwen-ai.

Skill directory

Use this skill's internal files to execute and learn. Load reference files on demand when the default path fails or you need details.

Location Purpose
scripts/tts.py Qwen TTS (HTTP API) — qwen3-tts-flash,

(truncated - see the full file via the links below)

File tree — 11 files
skills/audio/qianwen-audio-tts/SKILL.md
skills/audio/qianwen-audio-tts/references/agent-compatibility.md
skills/audio/qianwen-audio-tts/references/api-guide.md
skills/audio/qianwen-audio-tts/references/cosyvoice-guide.md
skills/audio/qianwen-audio-tts/references/execution-guide.md
skills/audio/qianwen-audio-tts/references/prompt-guide.md
skills/audio/qianwen-audio-tts/references/sources.md
skills/audio/qianwen-audio-tts/scripts/gossamer.py
skills/audio/qianwen-audio-tts/scripts/qianwen_lib.py
skills/audio/qianwen-audio-tts/scripts/tts.py
skills/audio/qianwen-audio-tts/scripts/tts_cosyvoice.py

Let your AI agent find skills like this

Example. Real query, live index.

You found this page by searching. An agent finds it by wishing: SkillFed indexes 56,283 agent skills by what they can do, searchable in plain language.

wish › “Convert text to speech using Qwen TTS models”

Give your agent the search over MCP, or paste the wish link into any chat. No install? Search from any chat →

Related skills

qwencloud-audio-tts
by QwenCloud · QwenCloud/qwencloud-ai

Turn written text into high-quality spoken audio using QwenCloud's TTS models. Choose between fast standard synthesis (qwen3-tts-flash) or instruction-guided style control (qwen3-tts-instruct-flash), or opt for premium quality via CosyVoice. Select from multiple voices and languages to match your content needs.

Apache-2.0updated Jul 2026
★ 34repo stars
qianwen-text
by QianWen-AI · QianWen-AI/qianwen-ai

qianwen-text lets you interact with Qwen language models for text generation, conversation, and code writing through an OpenAI-compatible interface. Choose from multiple Qwen models including qwen3.6-plus (recommended default), specialized code models, and reasoning variants. The skill handles API authentication, provides execution guides, and includes prompt engineering references.

Apache-2.0updated Jun 2026
★ 38repo stars
qianwen-ops-auth
by QianWen-AI · QianWen-AI/qianwen-ai

This skill walks you through QianWen API credential setup and validation. It handles both standard and Token Plan key types, manages environment variables and .env files, and provides verification steps to confirm your authentication is working correctly.

Apache-2.0updated Jun 2026
★ 38repo stars
qianwen-vision
by QianWen-AI · QianWen-AI/qianwen-ai

qianwen-vision lets you process images and videos through Qwen's vision models to extract text, understand visual content, and reason about complex scenes. Use it for OCR, chart/table analysis, multi-image comparison, and video comprehension across multiple model options tuned for speed, precision, or reasoning depth.

Apache-2.0updated Jun 2026
★ 38repo stars
qianwen-video-generation
by QianWen-AI · QianWen-AI/qianwen-ai

Generate videos through multiple input modes—text descriptions, single images, frame transitions, character role-play, or video editing—powered by Qianwen's Wan models. All operations run asynchronously; submit your request and poll for completion. The skill auto-detects your task and routes to the right model, with detailed reference guides for prompt engineering, polling patterns, and media workflows.

Apache-2.0updated Jun 2026
★ 38repo stars
Ttscn
by Agents365-ai · Agents365-ai/365-skills

Ttscn converts Chinese and multilingual text to natural speech across 11 cloud backends, from free Edge TTS to premium providers like ElevenLabs and OpenAI. Choose by use case—short video, audiobook, enterprise SSML, or lowest cost—with support for word-level timestamps, pause markers, and pronunciation overrides.

no license declared → metadata onlyupdated Jul 2026
★ 24repo stars

More skills qianwen-image-generation (Apache-2.0) · Tts (unlicensed) · Dashscope (AGPL-3.0) · Text To Speech (Unlicense) · aliyun-modelstudio-entry-test (MIT)

Tags
voice-synthesisaudio-generationmultilingual-ttsspeech-apivoiceover-toolemotion-controlstreaming-audioapi-integrationreal-time-speech