Audio | podcast | transcribe
Audio | podcast | transcribe
WHAT’S NEW
AI Deepfake Detector
Helps journalists screen video, audio, images, and text for possible AI generation or manipulation before publication. It combines C2PA provenance data, file metadata, and AI analysis to return a risk level with uncertainty notes rather than a definitive verdict; no signup is required, and each visitor receives free checks. DeepFakeCheck does not retain uploads after the request completes, while media sent to Google Gemini is subject to Google’s privacy policy. Two free checks, then a $6.99 monthly fee.
RJI Momentum Podcast
A series of conversations with news industry thought leaders about strategies that can help newsrooms navigate the ever-changing world of technology and audience behavior. In the first season of three episodes available now, RJI Executive Director Randy Picht gets three different newsroom perspectives about generative AI.
FlashEdit
A browser-based AI creative platform for generating and editing videos, images, audio and visual content. Its AI video generator lets users create short-form videos from prompts, images, videos, and other media references using multiple AI video models in one workflow. It is useful for journalists, educators, creators, and content teams who need quick social videos, explainers, story visuals, or media experiments without a complex editing setup.JOURNALISM BASICS
EconoFact
A non-partisan publication by Tufts University Fletcher School, bringing facts and incisive analysis to the national debate on economic and social policies.
Conversent
Turns your studio’s audio and video into knowledge that you control.
StudioEditor
A free, browser-based audio workspace in which recording, timeline editing, exporting, leveling and automatically ducking music under speech cost nothing. Designed for podcasters, creators, educators, interviewers, journalists and media, it combines a traditional audio editor interface with optional but unbeatable AI-tools, all in your browser.
Editor’s note: Before using any transcription tool, think about the security of your data. Read these articles from Politico and the Freedom of the Press Foundation about security issues with these tools and apps. For sensitive interviews, it might be best to transcribe it yourself.
Otter.ai
The Ferrari of transcription tools. Note privacy issues on free version, which give you 600 free minutes of transcription. Paid version has more privacy.
Descript
An AI-powered editor that automatically transcribes your audio and video recordings so that you can edit them just like text.
Whisper
Audio transcription tool.
Eleven Labs: Audio
Create sound effects and AI-generated voices with text-to-speech and other features.
Adobe Podcast: Enhance Speech
Free AI tool that cleans up your audio.
Alice Secure Recorder
“The world’s most confidential recorder and AI note-taker.” Used by journalists, healthcare and legal workers.
Speechify
Text-to-speech reader. Works with docs and PDFs, etc.
Murf.ai
Text-to-voice generator
ZenMic
Turn any text, blog, or RSS feed into a professional multi-speaker podcast. Custom voices, full script control and your own podcast RSS feed
Meta SAM Audio
AI audio editor dropped in late 2025 lets you drop out background noise to hear a person speaking, among many other features. Free download to your desktop.
Bark – Text to Audio AI Tool
Generates highly realistic, multilingual speech as well as other audio – including music, background noise and simple sound effects – all for free
Resound
Great for removing extra words.
Automat
Turn screen recordings or written instructions into production-ready automations for enterprise workflows
Turboscribe
Transcription tool known for accuracy
Uniscribe
Faster conversion of local audio and video files or YouTube videos to text using an optimized Whisper model. – Automatic generation of summaries, mind maps, and key Q&A. Supports exporting text content in various formats, such as .txt, .pdf, .docx, .srt, .vtt and .csv.
Waystars AI
Integrates AI tools for images and voice
FlowSpeech
A context-aware AI text-to-speech tool for creators, educators, marketers, and product teams. It supports emotion control, pause control, and more than 30 voices to generate more natural voiceovers, narration, and spoken-content assets. It is especially useful for product demos, tutorials, onboarding content, educational material, and other audio workflows.
Beepbooply
AI voice generator with more than 900 voices and 80 languages. Free model with paid subscriptions between $7 and $79.
Miso One
Advanced AI voice generator for hyper-realistic, expressive speech and one-shot voice cloning.
NeatScribe
Turns audio, video, recordings, and supported public video links into clean, readable transcripts. You can also generate subtitles, translate transcripts, and export results as TXT, DOCX, PDF, SRT, VTT, or LRC files. From meeting notes and lecture summaries to video captions and song lyrics, NeatScribe helps you save time on manual transcription work.
Spacebar.fm
A voice-to-text AI app that transcribes and transforms audio content. After recording with Spacebar, a “memory” is generated which provides a thorough recap of the content that’s been captured. Users are then able to explore new perspectives with AI –– chat with a single memory or multiple, recall important details, turn field notes into articles, podcasts, and more.
Findaway Voices by Spotify
Create AI audiobooks and publish to Spotify
Asyncflow v1.0
Text-to-speech AI with over 450 voice options
Mumble Note
AI voice note taker that turns ideas into actionable notes
Hume Octave
Create AI voices that understand emotional expression
Nieman Lab Prediction 2025: AI turns news into a conversation
Good Tape
Secure and automatic transcription
Scribe
ElevenLabs’ SOTA speech-to-text model
InfiniteTalk AI
Next-level conversational voice generation.
MP3 to Text
A lightweight, browser-based MP3 to text converter with TXT and SRT export. Upload an MP3, transcribe, then copy text, download TXT, or export SRT. You can start free with no credit card.
Octave TTS
Generate AI voices with emotional delivery
Y2Doc
Transform YouTube videos into structured documents
Lightricks
This AI platform offers an audio-to-video feature that lets you upload music, voice or sound effects to build a video.
AI Took My Voice. I Want it Back
Beatoven AI
AI composer for crafting the perfect background music
ScriptTimer
A free script timing calculator built for journalists, podcasters, and broadcast creators. Paste any script and instantly get your exact speaking time, word count, and per-paragraph timing breakdown — no sign-up required. It works in any language and supports slow, normal, and fast speaking speeds for radio segments, podcast episodes, and video packages.
Maibrain
Preserve the voice and experiences of your loved ones so you can interact with them in the future.
BeMusic AI: AI Music Generator
AI music generator built for anyone who wants to create original songs without a studio. Whether you need background music for a video, a full-length track for a podcast, or a singing photo for social media, BeMusic AI turns simple text descriptions into studio-quality music in under 30 seconds.
Veracity
A web app built for journalists that transforms interview recordings into a verifiable workspace, where every AI-generated sentence is traceable to the exact moment in the source audio. It combines transcription, summaries and quote extraction with a unique verification layer that lets users click any claim and hear the original source—or see it flagged if it can’t be confirmed. By checking both AI outputs and a journalist’s final article against source material, Veracity ensures accuracy, reduces errors, and protects credibility before publication.
Lyria 3
Google’s new AI music generation model in Gemini
Microsoft VibeVoice
Produces long-form, multi-speaker conversational audio, such as podcasts, from text. It addresses significant challenges in traditional Text-to-Speech (TTS) systems, particularly in scalability, speaker consistency, and natural turn-taking.
AI Translate Video
When you have local MP4/MP3 files that need translation Got a public video URL? Paste it—no download needed. Need separate subtitle files? We’ve got you covered. Want to translate your YouTube videos into multiple languages? Reach global audiences. Keep the original speaker’s voice? Voice cloning can help. Creating multilingual TikTok/Shorts videos to test different markets? Quick and easy.
WhisperTranscribe
Creates podcasts, blogs, and social content
Talktastic
Cleans up audio you record and summarizes it. Website app.
Udio
AI-music generator
Letterly
Captures your voice and lets AI tdo the writing. Phone app and desktop tool. Then tell it how to organize thoughts: Outline/summary for a class or a presentation.
Whisper Web
Free, browser-based transcription tool that converts audio and video files to text in 100+ languages with no signup required. It runs OpenAI’s Whisper model directly in your browser, so interview recordings never leave your device — a meaningful privacy protection for journalists. It’s a practical zero-cost option for reporters transcribing interviews on deadline.
Song Maker
Create complete music with AI. Use our AI song maker and AI song generator to make a song, make song ideas from lyrics to song.
StudioEditor
A free, browser-based audio workspace in which recording, timeline editing, exporting, leveling and automatically ducking music under speech cost nothing. Designed for podcasters, creators, educators, interviewers, journalists and media, it combines a traditional audio editor interface with optional but unbeatable AI-tools, all in your browser.
AI Deepfake Detector
Helps journalists screen video, audio, images, and text for possible AI generation or manipulation before publication. It combines C2PA provenance data, file metadata, and AI analysis to return a risk level with uncertainty notes rather than a definitive verdict; no signup is required, and each visitor receives free checks. DeepFakeCheck does not retain uploads after the request completes, while media sent to Google Gemini is subject to Google’s privacy policy. Two free checks, then a $6.99 monthly fee.
RJI Momentum Podcast
A series of conversations with news industry thought leaders about strategies that can help newsrooms navigate the ever-changing world of technology and audience behavior. In the first season of three episodes available now, RJI Executive Director Randy Picht gets three different newsroom perspectives about generative AI.
FlashEdit
A browser-based AI creative platform for generating and editing videos, images, audio and visual content. Its AI video generator lets users create short-form videos from prompts, images, videos, and other media references using multiple AI video models in one workflow. It is useful for journalists, educators, creators, and content teams who need quick social videos, explainers, story visuals, or media experiments without a complex editing setup.JOURNALISM BASICS
EconoFact
A non-partisan publication by Tufts University Fletcher School, bringing facts and incisive analysis to the national debate on economic and social policies.
Conversent
Turns your studio’s audio and video into knowledge that you control.
scryp
An AI transcription service from Vienna that turns interview, meeting and podcast recordings into accurate text with speaker labels. Files are encrypted in the browser before upload (AES-256-GCM) and transcription runs entirely in scryp’s own data centre in Austria, which makes it a GDPR-safe choice for source protection. It supports 134 languages, subtitle export (SRT/WebVTT) and AI summaries, with a 14-day free trial.
Voice to Notes
A powerful AI-powered tool that instantly converts your voice into clear, structured notes. It offers real-time transcription, automatic summaries, grammar correction, and smart formatting—all accessible across web, Android, and iOS. Designed for students, professionals, and creators, VoiceToNotes.ai helps you capture and organize your ideas effortlessly.
RJI Momentum Podcast
A series of conversations with news industry thought leaders about strategies that can help newsrooms navigate the ever-changing world of technology and audience behavior. In the first season of three episodes available now, RJI Executive Director Randy Picht gets three different newsroom perspectives about generative AI.
LegiTalk from CT Mirror
Tool that summarizes Connecticut legislative meetings and provides transcriptions for reporters to use. They can then click forward to the section of the video to confirm the transcript.
LyricsGenerator.io
Free AI song and lyrics creator
Sonix
Good speech to text tool for transcribing audio.
ChatGPT Advanced Voice
Now available on desktop apps for both Mac and Windows
A professional PDF to MP3 converter that turns any PDF into high-quality audio using advanced AI text-to-speech. You can convert PDFs to MP3 in seconds with 61 natural-sounding voices across 8 languages.
Command A Translate
Translation model
Speechmatics
Build voice-enabled apps with Speechmatics’ Startup Program and get up to $50k in credits to deploy to production*
ChatGPT Translate
Translate text, voice, or images across 50+ languages
AssemblyAI
Offers speech recognition and audio intelligence APIs for transcription, speaker diarization, and content moderation.
Nieman Lab: Local Newsrooms Are Using AI to Listen in on Public Meetings
ElevenMusic
Platform for AI song generation, remixing, creator payouts.
LyricsGenerator.io
AI music creation platform, combining a versatile Lyrics Generator with professional Song Production tools.
TranslateGemma
Google’s new family of open translation models
Attesta
An iOS recorder that announces out loud that it is recording, never trains on your audio, and leaves the file yours. The output is built to be usable, not just archived: top-tier models, speaker labels, and the action items, decisions, open questions and key people pulled out accurately. You can extend a recording with Continue, and refine or restyle the summary to match how you write.
Snappixify.net
Convert MP4/MP3, download Instagram videos, and more with our fast online tools. Supports H.264, 4K resolution, and all major formats.
Claude Cowork
Bring Claude Code’s agentic capabilities to everyday tasks
Livekit
New open-source turn detection model for natural voice AI.
Amazon Polly
Deploy high-quality, natural-sounding human voices in dozens of languages
Remento
AI biographer with Speech-to-Story technology that turns recorded interviews into a hardcover book of polished stories
Autopod
Save hours in your weekly production time in Adobe Premiere Pro with this pack of plug-ins. Free trial to get you started.
Amical
AI voice-to-text app for dictation, meetings and note-taking
GPT-4o-tts and Transcribe
OpenAI’s text-to-speech and speech-to-text tool
Restream.io
One livestream going to more than 30 audiences.
Wordcab One
Transcription, speech intelligence, and summarization, all in one tool.
Wondershare Virbo
A multiple languages video translator that can generate AI voice clones and lip syncs. It automatically generates subtitles, avatars scripts with AI.
MiniMax T2A-01 HD
Text-to-audio model enabling voice cloning with just 10 seconds of audio
AI Convert Hub
A fast, secure online file conversion and compression platform for video, audio, and images, supporting more than 200 formats with no software installation. It offers batch uploads, AI-optimized compression to preserve quality and advanced codec-aware controls for power users. With a free tier up to 1GB, multi-language support, and SEO-friendly positioning, it helps creators and teams quickly convert or compress files entirely in the browser.
AI Song Generator Free
Easy-to-use platform that creates high-quality, royalty-free music in various styles, perfect for videos, ads, and creative projects.
Notta Showcase
Translate videos into 15-plus languages while retaining the original voice.
InfiniteTalk AI
Create infinite‑length talking videos from any video or image. InfiniteTalk AI delivers razor‑accurate lip sync, expressive full‑body motion, and rock‑solid identity preservation—powered by next‑gen sparse‑frame technology.
Mumble Note
AI-powered voice notetaker designed for journalists and creators to instantly transform interviews, meetings, or idea sessions into clean, structured notes and action items. It automatically extracts key points, generates summaries, and tags content—so you can stay on story instead of scrambling for notes. With encrypted processing and support for images, links, and voice-to-text in 40+ languages, it blends privacy, versatility, and speed for every reporting workflow.
Tomedes AI Transcription
Supporting formats like MP3, MP4, WAV, and nearly 100 languages, it’s perfect for transcribing interviews, meetings, and lectures.
Stable Audio Open Small
Text-to-audio model for music sample
Nova-3 by Deepgram
New voice AI model for real-time multilingual transcriptions in real-world enterprise use cases
CloneDub
Convert audio into other languages using the same voices. Only audio files, YouTube, or audio links less than 15 minutes will work. Note: Many ethical concerns with this tool. Fact-check any translations for accuracy.
Poynter: AI Detection Tools for Audio Deepfakes Fall Short
Experts test four tools and show options on what to do.
Conversational AI
ElevenLabs tool that allows users to seamlessly add voice capabilities in 31 languages
Clipto
Convert audio to text in 99 languages.
TalkText
An AI-powered dictation tool that transforms speech into text
1B CSM
Text and audio-to-speech model
SynthID
Audio tool from Google Deepmind
Resemble.ai
Deepfake audio detection
Transkriptor
Convert audio to text quickly
Pindrop
Realtime audio deepfake detection
Tad.ai
Create original music from prompts
Voiser AI
Transcribe, summarize, and translate videos and recordings
Hume AI Voice Control
Helps developers create consistent, custom AI voices by adjusting 10 sliders and settings.
AgentPlace
Create AI-driven websites and apps through simple text instructions
DiffRhythm
Generate complete 4-min songs w/ vocals in just 10 seconds
Akool.com
Create video, audio and characters. Also translate audio into other languages
Adobe Enhance Speech
Upload any audio recording with background noise and immediately get a clean version.
Video to Text
Video to text is an ai-powered transcription service that converts video and audio files into clean, exportable text. the product is designed for creators, teams, and individuals who need fast, accurate speech-to-text conversion without setting up their own transcription pipeline.
Read PDF Aloud
Let AI Read Your PDF Aloud with Natural Voice. Free AI-powered PDF speaker. Listen to any PDF document with natural-sounding voices in 140+ languages.Read your PDF aloud with one click.
Good Tape
A security-first AI-powered transcription tool created by journalists and known for its fast turnaround and unprecedented accuracy, when it comes to audio and video transcriptions. Trusted by over 2.5 million users worldwide, GoodTape has transcribed more than 3 million hours of audio and saved 12 million hours of work globally. Good Tape understands the importance of confidentiality when working with sensitive sources and materials and does not use the customer’s transcription files for AI learning of any kind.
ChatGPT Record
Capture, summarize, and transcribe audio with ChatGPT
Magenta RealTime
Google’s new open-weighted live music model
Sonix
Powerful transcription tool. 30-minute free limit and support more than 40 languages.
Fireflies
Its free plan offers unlimited transcription, with storage limits.
All Voice Lab
Delivers AI-powered voice cloning capabilities
ACE Studio
AI workstation to generate studio-quality singing vocals
Projects by ElevenLabs
Build long-form audio
Stock Music GPT
Instant royalty-free stock music, sound effects and song covers, generated by AI.
Workflowy
Powerful outliner helps to jot down ideas, thoughts, writing topics and block drafts and to easily re-arrange and organize them. Can also be used to manage tasks and projects. Intuitive. Fast. Definitely my favorite outliner. Free version (all features with monthly limit). Pro version $8.99 a month.
MusicFX
Google’s free text-to-song creation tool
Orpheus TTS
Open-source text-to-speech AI with natural emotion
Talo
Real-time voice translator for video calls
GitPodcast
Build podcasts to understand GitHub repositories
Newsroom Robots Podcast: Congresswoman Anna Eshoo on Shaping the Future of AI Policy
Peech
Effortlessly transforms any text into incredibly realistic AI-generated audio. Peech supports over 50 languages, including English, French, German, Italian, Spanish, and more.
Listnr
A free voice AND video generator in multiple languages
Papercup
AI video and audio dubbing tool
MusicFX
Google tool that lets you describe what you want and it will write a song.
Eddy
Free transcription tool from Headliner app.
Speech-02
Minimax’s text-to-speech AI supporting over 30 languages
Vapi
Build voice AI agents
Everlit Audio
Audio creation tool. It has a WordPress plug-in, embeddable player, APIs, voice replication and more.
Eleven v3
SOTA text-to-speech model with support for 70+ languages
Bland TTS
Voice AI with enhanced control
HunyuanVideo-Avatar
Multi-character talking videos from audio
Musick AI
Creating high-quality, original music across various genres. user-friendly interface, time-saving capabilities, AI filters, mood templates, and interactive delivery.
AudioRead
Use AI to listen to articles, PDFs, emails, etc. in your podcast player. Read while walking, driving, cleaning and more.”Read” while walking, driving, cleaning, and more
Superwhisper
Fast and accurate voice to text
Udio Audio Inpainting
Select a portion of an AI-generated music track and regenerate it. Be careful with music rights using this tool.
Stability AI’s Stable Audio Open
Generates up to 47-second audio samples based on text descriptions. It’s trained on thousands of royalty-free music samples.
Transcript LOL
Transcribes podcasts, videos and meetings
GitPodcast
Generate engaging podcasts to understand GitHub repositories
Supertone Shift
A real-time voice changer
ElevenLabs Audio Native
Add narration to your blog or news site
Adobe Express Animate from Audio
Upload any audio and animate characters with AI
AI Cover Art
A streamlined platform that uses artificial intelligence to create high-quality music covers quickly. The tool transforms songs with AI-generated vocals that sound natural and professional. Users can select from different voice styles, adjust vocal parameters, and export their creations in minutes. Perfect for musicians, content creators, and music enthusiasts who want to explore new versions of their favorite songs without complex equipment or technical skills.
YouTube Transcription Generator
A free, no‑login tool for instantly generating transcripts from any YouTube video. Users can extract subtitles in multiple languages, search within the transcript, and copy or download TXT, SRT, or VTT files, with or without timestamps. Designed for creators, students, researchers, and editors, it streamlines workflows by making it fast and distraction‑free to grab accurate transcripts and jump to exact moments in a video.
Tomedes AI Transcription
A free, browser-based speech-to-text tool that converts audio files into text using multiple leading AI engines—Whisper, Gemini, and Amazon Transcribe. It generates up to three transcription results per file, allowing users to compare and select the most accurate version. With no sign-up, no ads, and multilingual support, it’s ideal for journalists, podcasters, and researchers.
SeedanceHub | AI Video Generator
Transforms text prompts or images into high-quality videos. It offers smooth motion, vivid visuals, and fast generation speeds.
Chuchotis
A whispered-based application on MacOS (Apple silicon) for 100% confidential voice-to-speech transcription, developed by a french journalist, and heavily tested by his colleagues. It offers many tools to improve transcription (prompt, hallucinations & incomplete segment, batch processing, tools for edition… Localized in English, french and spanish, works on 100 languages.
EVI 3
Create custom voices through speech-to-speech interactions
Conversational AI 2.0
ElevenLab’s updated AI voice agent platform
PlayDiffusion
PlayAI’s open-source audio inpainting model
Cassette AI
Music-generation tool.
Pdf2audio
Open-source tool that converts PDFs to podcasts, lectures, summaries and more
Minutes
Automate meeting notes and transcriptions with AI
Alphana
Convert podcasts and video into several short, viral pieces.
Eleven Labs: Multispeaker Podcasts
Works like Notebook LM
Muzix
Transform text into music with AI music generator. Create custom songs and instrumental tracks in minutes, no musical experience needed.
Plus AI
Google Slides tool generates presentations from any text input.
SOUNDRAW
Uses rights-free audio to build music.
OpenMusic
Create custom tunes from text descriptions
Udio Lyric Editor
Create and refine song lyrics based on melody
Soundraw
AI music generator
Mubert
Generative AI music for videos and podcasts
Artist.io
Create voiceovers for video using AI
VoiceType
Most professionals spend 5-10 hours a week typing. But this text-to-voice tool speeds that up.
Songscription
Turning audio clips instantly into sheet music
Suno
Generate a song from any sound
Voice Isolator by Eleven Labs
Strip background noise from interviews, films and podcasts
SpeakHints
Real-time AI speech copilot for any situation
Minvo
AI video shorts platform for podcasts and more.
Wondercraft AI Audio
Creates hyper-realistic audio ads
Soundraw
Generate music/beats using AI.
Jotify
Turn research into stories and audio
Mixhalo
Audio translation on your phone
ElevenLabs Text to Sound Effects
Generate any sound from a prompt
Pika Labs
Add sound effects to your videos
Synthflow AI
Create conversational AI voice agents without coding.
Auphonic
AI sound engineer for podcasts, etc.
MacWhisper Pro
Free transcription tool with strong privacy
Jellypod
Transform your inbox into a personalized daily podcast
izTalk
Translation and speech recognition
OpenVoice
A free tool to clone voices
Letterly
Record your voice and let AI turn it into well-written text
LALAL AI
Remove voices and instrumental audio splitter
Deepgram Aura
A text-to-speech API built for real-time conversations.
CleanVoice
Cleanvoice is an artificial intelligence which removes filler sounds, stuttering and mouth sounds from your podcast or audio recording
Coval
Make voice and chat agents with seamless simulation
PopPop AI Vocal Remover
Free AI-powered tool to separate or remove vocals from any song
ArticleReader
Turns text into captivating audio
EVI-2 by Hume
AI speech with emotion, using several voices and personalities
CloneDub
Automatically dubs your videos and podcasts
Dubformer
AI-powered dubbing solution with broadcast quality
Video Insights
Video and audio summarization and transcription
Dubformer
AI audio dubbing tool
MixAudio
AI music generator and editor
Poynter: How Text-to-Speech Technology Can Help Journalists Avoid Copy Errors
Udio New Features
Generate AI music longer than 2 minutes and extend tracks up to 15 minutes
Hive AI Detection
Check photos, video, audio and text. Freemium tool and you can request a demo.
PocketPod
Creates personalized podcasts that match your interests.
Replica Voice AI
Ethical voice AI for creators and business
Chuchotis
Chuchotis (French word for whisper) is a fee-based speech-to-text application based on OpenAI’s whisper. It runs fully locally on any M-silicon-based Mac, to preserve privacy of recordings and transcriptions.
FlowTunes
Endless AI-generated music for focused work
SpeechTexter
Type with your voice
Beatoven AI
Create royalty-free background music.
SpeakAI
Get transcription, research, data analysis and NLP software.
Udio Audio Prompting
Upload a sound and let AI generate a song
AI Jukebox
In-browser text-to-music generation
Soundry AI
AI sound sample VST for music creation and DJing
Ramble Fix
An audio to text tool. Speak your mind and it will paraphrase/rewrite what you say. Five free uses per month, then $10 monthly for more recordings.
Teachable Machine
Teachable Machine is a web-based Google tool that makes creating machine learning models fast, easy, and accessible to everyone. Train a computer to recognize your own images, sounds, and poses. A fast, easy way to create machine learning models for your sites, apps, and more – no expertise or coding required.
Adobe Speech Enhancer
Speech enhancement makes voice recordings sound as if they were recorded in a professional studio. Be careful with ethics and edits with this tool.
Google AudiopaLM
A large language model for speech understanding and generation. AudioPaLM fuses text-based and speech-based language models, PaLM-2 and AudioLM [Borsos into a unified multimodal architecture that can process and generate text and speech with applications including speech recognition and speech-to-speech translation.
Easy_Peasy.ai
Create images with text, transcribe audio and create audio with this Swiss Army knife of an AI tool
Rizzle AI
Convert text and podcasts into captivating videos
Toasty AI
AI content creation for podcasts
CastMagic
Turn audio into content
Audio Notes AI
Organize thoughts into structured notes with AI
Article Audio
Convert your article to audio. Free model with paid upgrade.
Pickle
Lifelike AI clones lip-syncing to your voice in real-time. Be careful with ethics if you are using this tool in a journalistic way.
Wordtune
Rewrite your thoughts in different styles.
Castmagic
Turn long-form audio into ready-to-use content assets, instantly. Requires a log-in to PartnerSnack first.
Co-Producer Pack Generator
Generate authentic music sample packs
Music Maestro GPT
A ChatGPT GPT assistant for music creation
MusicFX DJ
Google has added ‘DJ Mode’ in MusicFX, the generative text-to-music tool powered by Google’s MusicLM
PDFToMP3
Transforms PDFs into MP3 formats
SwellAI
AI assistant for podcast producers
Nendo
Open-sourced AI music generation models
Musicgen-remixer
Remix music into other styles
Respeecher
Hollywood-grade AI voice generation
OpenAI Advanced Voice
Creat conversations with AI, available forChatGPT Plus subscribers
Whisper WebGPU
A real-time in-browser speech recognition with OpenAI Whisper. The model runs fully on-device and supports transcription across 100 different languages.
MAGNeT
A high-quality text to sound and text-to-music
Podsqueeze 2.0
Podcast content repurposing
Hypernatural
Use it to generate audiogramshttps://www.podmind.ai/
Dub AI
AI-powered voice cloning and translation
Read This AI
Transform text into high-quality audio effortlessly
Voicenotes
Notetaker that transcribes and provides info on your thoughts
izTalk
Overcome language barriers with instant AI translation
Text Reader AI
Convert text to speech with free, realistic AI voices
BlueDotHQ
AI meeting recorder backed by Google
AudioPen
Just hit record. Then start talking. AudioPen will clean things up when you’re done.
TalkNotes
Transcribes your voice notes in more than 50 languages
Lalal
Extract vocal, accompaniment and various instruments from any audio and video. One-time fees start at $15
Zealous
Convert audio files into social media posts. Tool is marketing-driven but has social media desk benefits. Free with reasonable upgrades.
Voicemod
An AI voice creator that lets users make their voices by changing gender, age and tone.
Mubert
Create soundtracks for your projects with AI. Freemium account
Sumly
Audio and podcast summary tool
Director Mode by Wondercraft
Fine-tune and direct AI voices through prompts
Audeus
Text-to-speech AI for PDFs, email, etc.
Voicify
Create AI covers using AI in seconds with Voicify, with hundreds of community-uploaded AI voice models available for creative use now.
Krisp
Noise-canceling app
Altered
Voice-changing software. Be sure to use this ethically and carefully.
VoiceMod
Free real-time voice-changer. Be sure to use this ethically and carefully.
BeyondWords
Text-to-speech tool. Free account with paid upgrades.
Samplab
Edit audio with AI
Vocal Remover
Separate voice from music out of a song free with powerful AI algorithms
AI Jingle Maker
Create jingles, radio sweepers, podcast intros, audio promos and more.
AssemblyAI | Playground
Production-ready AI models for speech recognition, speaker detection, audio summarization, and more through our API. Quickly test below using any YouTube link, audio file or video file.
Narakeet
Text-to-speech tool for creating voiceovers. Pay by the length of the audio.
Cassette.ai
Music generation and text-to-audio tool
CastMagic
Upload and .mp3 and download a transcript.
Wavel
AI video dubbing with realistic voices, subtitles, accents, etc..
Flowjin
AI-generated short video clips for audio or video podcasts
Songburst
Mobile AI music generation app with prompt enhancer
Byrdhouse AI 2.0
Multilingual video call AI interpreter
Listnr
Free AI voice and video generator, choose from 900+ voices in 142 languages. Get started for free, download in MP4/MP3/WAV formats.
Listnr
Text-to-voice and text-to-video tool.
Assembly
Exposes AI models for speech recognition, speaker detection, speech summarization, and more.
Khanmigo
AI teaching assistant from Khan Academy
Poised
AI-powered communication coach that helps you speak with confidence and clarity. Private and secure, an essential tool for digital-first workplaces. Free trial with personalized plans.
Kits.ai
Create AI voices or modify your own with a library of commercial use and officially licensed artist voices.
BeyondWords
Convert writing into audio. Free for up to 30,000 words, then paid.
Notta
Notta is an AI-based voice-to-text transcription software that supports 104 languages. Notta excels at transcribing and summarizing audio or video files, online meetings, and voice recordings. Additionally, Notta offers a suite of team alignment features, enabling users to schedule and transcribe Zoom, Google Meet, and Teams meetings, among other functionalities.
Boomy – Make Generative Music with Artificial Intelligence
Whisper
Popular AI tool for transcription.
Fathom
Note-taking and transcription tool for Zoom
Media.io
A collection of AI tools ranging from video/image enhancement to audio.
Wellsaid
Creates texts from voiceovers in seconds
Transistor
Quick and precise speech-to-text conversion for podcasts
PolyAI Pheme
Generate conversational voices for phone-call apps
Nolan Free Script-Writing Software
Decoherence
Make AI music videos
Castmagic
Create podcast notes by uploading an .mp3 file download all your post-production content.
Kaiber
An AI creative lab on a mission to unlock creativity through powerful and intuitive generative audio and video. Upload a song, add a touch of your artistic style, and let its audio analysis technology do the visual work. Monthly plans ranging from $5 to $25 with a free trial.
GIJN: Podcasts for Investigative Journalists
Global podcasts in a variety of languages
Suno.ai
Create songs
AI Phone
Provides live transcription and teal-time translation during phone calls using AI.
TurboScribe
Converts all audio and video files into transcripts
Songburst AI
AI-driven music generator
Papercup
AI dubbing for video
Respeecher
Clones voices for content creators
Bloks
Free desktop note-taking app
Wavel.ai
Dub videos in more than 30 languages. Free with paid upgrades. Be careful with ethics with this tool.
AI PODCASTING TOOLS
Adobe Podcast
AI audio recording and editing, all on the web.
Podcastle
A cloud recorder and AI-powered editor that lets you record a remote interview, edit and mix
Podcraftr
Repurpose blog posts, email newsletters, training materials into podcasts.
Wondercraft
Streamline your podcast with only text.
Auphonic
Upload your file, and Auphonic will automatically optimize your recording and clean it.
Castup
A professional audio and video editing service for podcasters and podcast networks. Castup offers subscriptions starting at $30 per episode, or 40 cents per published minute. Castup also offers a ChatGPT-powered podcast assistant called Castup AI, which can help you record and promote episodes.
Streamyard
A professional live streaming and recording studio in your browser. Record your content, or stream live to Facebook, YouTube, and other platforms. Freemium account with paid models starting at $20 a month.
Audioread: Read. In Audio.
Turns articles, PDFs, etc. into podcasts using this AI-driven tool
Cleanvoice.ai
AI audio tool that can remove unwanted sounds and speech imperfections.
Listener.fm
Use AI to generate podcast titles, descriptions, and show notes in seconds
Promptcast
Summarizes any podcast with AI
Listnr
Paste text into the text-to-speech converter, and the app will convert it into one of their 600 voices.
Alitu
AI-powered one-stop shop for podcast production. Record remotely and edit.
Podwise
A learning app for podcast listeners
Newsroom Robots: Podcast About AI
Google’s Music LM
Create your own music in Google’s test kitchen.
Mubert
AI-generated music
The Essential AI Toolkit for Journalists and Content Creators
All the tools listed are free to use and open-source. Audio, images, web scraping and other tools
Boomy
AI music generator. Free version with paid models ranging $10 to $30 a month.
Listen Notes
A ChatGPT plug-in that allows you to search for podcasts by person or topic.

Notify me when this page changes
Resources | Tools | Training