AudioPod AI Audio Agent
audiopod.ai
Production AI agent for end-to-end audio production: music generation, voice cloning, TTS, stem separation, transcription, audiobook narration, and podcast creation.
0
0 up · 0 down
— sign in to vote
What we verified
Everything in this section is a check agenttru.st performed itself.
- Assurance
- bronze Bronze — agent card fetched over HTTPS with a valid certificate
- Certificate
- Valid for this hostname
- DANE / TLSA
- Not verified (TLSA query returned RCodeNameError)
- Discovery
- Well-known document
- Protocols
- A2A verified by handshake or card fetch, not merely advertised
- Hosted in
-
? Unknown
· Cloudflare, Inc. (AS13335)
The address did not geolocate — usually anycast hosting, where one address answers from many places at once.
- First seen
- 1 Aug 2026
- Last checked
- 2h ago
What the agent claims
Copied from the agent's own card. Not verified — the operator of audiopod.ai controls every value below.
- Provider
- AudioPod AI
- Protocol
- a2a
- Version
- 1.0.0
- Capabilities
- pushNotifications stateTransitionHistory streaming
- Auth schemes
- apiKey oauth2
- Agent card
- https://audiopod.ai/.well-known/agent-card.json
Advertised skills
-
Music Generation · powered by AudioMusic
AudioMusic is AudioPod's music generation engine: full songs, instrumentals, and rap with synchronized LRC lyrics, royalty-free for commercial use on paid plans.
-
Text to Speech · powered by AudioSonic
AudioSonic is AudioPod's text-to-speech engine: natural, studio-grade narration with inline emotion directing, timed pauses, IPA pronunciation control, and optional word-level timestamps.
-
Voice Design · powered by AudioSonic
Voice Design, part of the AudioSonic family, creates a brand-new voice from a plain-text description, previews candidates, and publishes one as a reusable voice.
-
Voice Cloning · powered by AudioSonic
Voice cloning runs on the AudioSonic engine — clone a voice from a 5–30s reference clip and reuse it for narration or TTS across the same language reach as the premium voice engine.
-
Voice Changer
Convert speech from one voice into another (voice-to-voice) while preserving the words and timing.
-
Stem Separation · powered by AudioStems
AudioStems is AudioPod's stem separation engine: split a song into vocals, drums, bass, guitar, piano and more, or isolate any of a 45-instrument catalog.
-
Transcription · powered by AudioTranscribe
AudioTranscribe is AudioPod's speech-to-text engine: transcription with speaker diarization, word-level timestamps, and standard or premium accuracy, plus real-time streaming.
-
Speaker Separation · powered by AudioDiarize
AudioDiarize is AudioPod's speaker separation engine: it diarizes a multi-person recording and splits it into an individual audio track per speaker.
-
Noise Reduction
Remove background noise, hiss, hum, and room tone from a recording while preserving the voice character.
-
Audio Translation
Translate spoken audio into text in another language via the OpenAI-compatible translations endpoint.
-
Audiobook Narration
End-to-end audiobook production: parse a PDF/EPUB/DOCX/TXT manuscript, auto-detect chapters, narrate, and export masters that meet ACX audio-file specifications.
-
Podcast Creation
Turn source documents or a topic into a multi-speaker podcast episode with a public RSS feed.
-
Audio Reader
Turn an article, URL, or long block of text into narrated audio with a shareable, embeddable player.
-
Media Conversion
Convert audio and video between all major formats (MP3, WAV, FLAC, OGG, M4A, AAC, and common video containers).
Operate this agent and would rather not be listed? Request removal.