← Back to Skills Marketplace
cellcog

Audio Generation

by CellCog · GitHub ↗ · v1.0.17 · MIT-0
darwinlinuxwindows ✓ Security Clean
7527
Downloads
4
Stars
0
Active Installs
18
Versions
Install in OpenClaw
/install audio-generation-cellcog
Description
AI audio generation and text-to-speech powered by CellCog. Voiceover, narration, voice cloning, avatar voices, sound effects, music, podcasts, dialogue. Three voice providers (OpenAI, ElevenLabs, MiniMax). Professional audio production from text prompts.
Usage Guidance
Before installing, confirm you are comfortable sending audio prompts and any voice samples to CellCog and its providers. Use cloned voices only for yourself or with clear permission from the voice owner, and check CellCog's retention and deletion options for uploaded samples and avatar voices.
Capability Assessment
Credentials
Network/API use and uploading prompts or voice samples to CellCog and underlying providers are proportionate to audio generation and disclosed by the workflow. This does involve externally processed personal or creative content.
Install Mechanism
The skill declares a Python dependency on cellcog, requires CELLCOG_API_KEY, and suggests npx, openclaw, plugin, or pip setup paths. There are no executable files in the artifact, but users should review the referenced CellCog dependency and setup flow before use.
Instruction Scope
The instructions show user-directed CellCog SDK calls and provider selection guidance. I found no hidden role changes, prompt injection, destructive commands, or unrelated data-access instructions.
Persistence & Privilege
The artifact does not request local persistence, background execution, privilege escalation, or broad filesystem access. Avatar voice samples and cloned voice IDs may persist in the user's CellCog account, so users should understand retention and deletion controls.
Purpose & Capability
The stated purpose is AI audio generation, text-to-speech, sound effects, music, and avatar voices; the capabilities described in SKILL.md fit that purpose. Voice cloning is sensitive, but it is clearly presented as part of the skill and framed around the user's own voice.
How to Use
  1. Make sure OpenClaw is installed (local or Docker)
  2. Run the install command in chat: /install audio-generation-cellcog
  3. After installation, invoke the skill by name or use /audio-generation-cellcog
  4. Provide required inputs per the skill's parameter spec and get structured output
Version History
v1.0.17
Content updated.
v1.0.16
Content updated.
v1.0.15
Content updated.
v1.0.14
Display title updated.
v1.0.13
- Skill name updated from "audio-cog" to "audio-generation-cellcog" to clarify functionality. - Updated documentation in SKILL.md for improved clarity and guidance. - Removed redundant skill-card.md file for simpler distribution. - No changes to core functionality or usage; documentation and naming updates only.
v1.0.12
- Added environment requirements for the skill: now specifies needed binaries (python3) and environment variables (CELLCOG_API_KEY) in metadata. - No user-facing feature or functionality changes; documentation in SKILL.md updated to reflect setup prerequisites.
v1.0.11
- Updated documentation for clarity and accuracy in SKILL.md - Improved and expanded description, highlighting support for podcasts and dialogue - Clarified agent usage: now specifies “all agents except OpenClaw” for blocking chat integration example - Refined instructions and language throughout for easier onboarding and provider selection - No code or functional changes—documentation update only
v1.0.10
- Improved documentation in SKILL.md with more concise descriptions and clearer usage instructions. - Expanded SDK usage code sample to explicitly show client initialization. - Updated skill description to highlight avatar voices and music generation up to 10 minutes. - Enhanced formatting and clarified steps for using voice providers, avatars, sound effects, and music features.
v1.0.9
- Simplified and clarified the description to emphasize major features and use cases. - Reorganized usage instructions for easier onboarding, highlighting SDK references up front. - Added new "If CellCog is not installed" section with agent-specific installation guidance. - Streamlined sections for voice providers and capabilities; removed duplicative wording. - Made instructions for different agent types (OpenClaw, Cursor, etc.) more concise and highlighted code output. - Shortened and focused provider feature explanations for improved readability.
v1.0.8
- Expanded SKILL.md with detailed guidance on provider selection (OpenAI, ElevenLabs, MiniMax) and their strengths. - Added new tables outlining voice provider scenarios, voice options, customization tips, and emotion tag usage. - Provided explicit examples for avatar/cloned voices and usage of custom avatars for personalized narration. - Enhanced sections on sound effects and music generation, including sample prompts and best practice tips. - Clarified multi-language support and practical agent usage for audio generation. - Included a new "Tips for Better Audio" section to help users optimize results with the skill.
v1.0.7
- Expanded description and usage details for TTS, music, SFX, and podcast production features - Clarified role of supported providers: OpenAI, ElevenLabs, and MiniMax, with highlights for multi-voice and avatar voices - Documented new features: multi-voice dialogue, podcast pipeline, 160+ voices, and output in MP3/WAV - Updated related skills section for audio, music, podcast, and video generation - Refined and reorganized documentation for easier discovery of key capabilities
v1.0.6
audio-cog 1.0.6 - Added explicit Python code examples for agent-based and blocking audio generation using the SDK. - Clarified that OpenClaw agent mode is recommended for long tasks and provided notify_session_key details. - Referenced the main cellcog skill for advanced SDK usage, delivery modes, and file handling. - Documentation updates only; no functional or interface changes.
v1.0.5
- Added OS compatibility metadata for Darwin, Linux, and Windows. - Updated skill description for improved clarity and SEO. - Added homepage link to CellCog website. - Improved formatting and metadata structure in SKILL.md. - No changes to core functionality; documentation and metadata only.
v1.0.4
- Adds support for three voice providers: OpenAI, ElevenLabs, and MiniMax, each with unique capabilities. - Introduces avatar/cloned voice generation via MiniMax, allowing users to create audio in their own voice. - Expands features to include standalone sound effects (up to 30 seconds) and longer music generation (up to 10 minutes). - Clarifies provider recommendations by scenario, with detailed guidance on emotional tags (ElevenLabs) and fine-grained controls (MiniMax). - Updates documentation with new usage examples, usage tips, and multi-language support across all providers.
v1.0.3
- Added author and dependencies fields to SKILL.md for clearer metadata. - Updated prerequisite instructions to refer to the cellcog skill by name. - Minor edits for clarity and consistency in setup and usage guidance.
v1.0.2
audio-cog 1.0.2 - Added detailed documentation of all 8 available CellCog voices, including usage recommendations and voice characteristics. - Included guidance on choosing voices by content type and how to customize styles (accent, emotion, pacing, etc.). - Added music licensing statement: all generated music is royalty-free and usable for any commercial purpose. - Expanded multi-language support list and updated multi-language example prompts. - Updated example prompts and tips to reflect voice selection and new usage patterns.
v1.0.1
- Adds metadata (including an emoji) for enhanced identification. - Updates quick-start usage pattern to use `create_chat` with simplified, fire-and-forget execution and notification model (v1.0+). - Clearly standardizes on `chat_mode="agent"` as optimal for all audio tasks, deprecating previous agent team recommendations. - Updates guidance to reflect new workflow and best practices. - No change in feature set or audio capabilities.
v1.0.0
- Initial release of audio-cog: Professional AI audio generation powered by CellCog. - Supports text-to-speech, voice synthesis, narration, voiceovers, podcast production, music creation, and sound design. - Offers voice customization (gender, age, emotion, accent, pacing, tone). - Enables music and background audio generation with detailed control (genre, tempo, mood, instruments, duration). - Multi-language speech generation and various audio output formats. - Includes prompt examples, guidance for agent team mode, and detailed usage tips.
Metadata
Slug audio-generation-cellcog
Version 1.0.17
License MIT-0
All-time Installs 0
Active Installs 0
Total Versions 18
Frequently Asked Questions

What is Audio Generation?

AI audio generation and text-to-speech powered by CellCog. Voiceover, narration, voice cloning, avatar voices, sound effects, music, podcasts, dialogue. Three voice providers (OpenAI, ElevenLabs, MiniMax). Professional audio production from text prompts. It is an AI Agent Skill for Claude Code / OpenClaw, with 7527 downloads so far.

How do I install Audio Generation?

Run "/install audio-generation-cellcog" in the OpenClaw or Claude Code chat to install it in one step — no extra setup required.

Is Audio Generation free?

Yes, Audio Generation is completely free, licensed under MIT-0. You can download, install and use it at no cost.

Which platforms does Audio Generation support?

Audio Generation is cross-platform and runs anywhere OpenClaw / Claude Code is available (darwin, linux, windows).

Who created Audio Generation?

It is built and maintained by CellCog (@cellcog); the current version is v1.0.17.

💬 Comments