Audio | podcast | transcribe

Audio | podcast | transcribe

Notify me when this page changes
Powered by visualping

WHAT’S NEW

AI Deepfake Detector
Helps journalists screen video, audio, images, and text for possible AI generation or manipulation before publication. It combines C2PA provenance data, file metadata, and AI analysis to return a risk level with uncertainty notes rather than a definitive verdict; no signup is required, and each visitor receives free checks. DeepFakeCheck does not retain uploads after the request completes, while media sent to Google Gemini is subject to Google’s privacy policy. Two free checks, then a $6.99 monthly fee.

RJI Momentum Podcast
A series of conversations with news industry thought leaders about strategies that can help newsrooms navigate the ever-changing world of technology and audience behavior. In the first season of three episodes available now, RJI Executive Director Randy Picht gets three different newsroom perspectives about generative AI.

FlashEdit
A browser-based AI creative platform for generating and editing videos, images, audio and visual content. Its AI video generator lets users create short-form videos from prompts, images, videos, and other media references using multiple AI video models in one workflow. It is useful for journalists, educators, creators, and content teams who need quick social videos, explainers, story visuals, or media experiments without a complex editing setup.JOURNALISM BASICS

EconoFact
A non-partisan publication by Tufts University Fletcher School, bringing facts and incisive analysis to the national debate on economic and social policies.

Conversent
Turns your studio’s audio and video into knowledge that you control.

StudioEditor
A free, browser-based audio workspace in which recording, timeline editing, exporting, leveling and automatically ducking music under speech cost nothing. Designed for podcasters, creators, educators, interviewers, journalists and media, it combines a traditional audio editor interface with optional but unbeatable AI-tools, all in your browser.



Editor’s note: Before using any transcription tool, think about the security of your data. Read these articles from Politico and the Freedom of the Press Foundation about security issues with these tools and apps. For sensitive interviews, it might be best to transcribe it yourself.

Otter.ai
The Ferrari of transcription tools. Note privacy issues on free version, which give you 600 free minutes of transcription. Paid version has more privacy.

Descript
An AI-powered editor that automatically transcribes your audio and video recordings so that you can edit them just like text. 

Whisper
Audio transcription tool.

Eleven Labs: Audio
Create sound effects and AI-generated voices with text-to-speech and other features.

Adobe Podcast: Enhance Speech
Free AI tool that cleans up your audio.

Alice Secure Recorder
The world’s most confidential recorder and AI note-taker.”  Used by journalists, healthcare and legal workers.

Speechify
Text-to-speech reader. Works with docs and PDFs, etc.

Murf.ai
Text-to-voice generator

ZenMic
Turn any text, blog, or RSS feed into a professional multi-speaker podcast. Custom voices, full script control and your own podcast RSS feed

Meta SAM Audio
AI audio editor dropped in late 2025 lets you drop out background noise to hear a person speaking, among many other features. Free download to your desktop.

Bark – Text to Audio AI Tool
Generates highly realistic, multilingual speech as well as other audio – including music, background noise and simple sound effects – all for free

Resound
Great for removing extra words.

Automat
Turn screen recordings or written instructions into production-ready automations for enterprise workflows

Turboscribe
Transcription tool known for accuracy

Uniscribe
Faster conversion of local audio and video files or YouTube videos to text using an optimized Whisper model. – Automatic generation of summaries, mind maps, and key Q&A. Supports exporting text content in various formats, such as .txt, .pdf, .docx, .srt, .vtt and .csv.

Waystars AI
Integrates AI tools for images and voice

FlowSpeech
A context-aware AI text-to-speech tool for creators, educators, marketers, and product teams. It supports emotion control, pause control, and more than 30 voices to generate more natural voiceovers, narration, and spoken-content assets. It is especially useful for product demos, tutorials, onboarding content, educational material, and other audio workflows.

Beepbooply
AI voice generator with more than 900 voices and 80 languages. Free model with paid subscriptions between $7 and $79.

Miso One
Advanced AI voice generator for hyper-realistic, expressive speech and one-shot voice cloning.

NeatScribe
Turns audio, video, recordings, and supported public video links into clean, readable transcripts. You can also generate subtitles, translate transcripts, and export results as TXT, DOCX, PDF, SRT, VTT, or LRC files. From meeting notes and lecture summaries to video captions and song lyrics, NeatScribe helps you save time on manual transcription work.

Spacebar.fm
A voice-to-text AI app that transcribes and  transforms audio content. After recording with Spacebar, a “memory” is generated which provides a thorough recap of the content that’s been captured. Users are then able to explore new perspectives with AI –– chat with a single memory or multiple, recall important details, turn field notes into articles, podcasts, and more.

Findaway Voices by Spotify
Create AI audiobooks and publish to Spotify

Asyncflow v1.0
Text-to-speech AI with over 450 voice options

Mumble Note
AI voice note taker that turns ideas into actionable notes

Hume Octave
Create AI voices that understand emotional expression

Nieman Lab Prediction 2025: AI turns news into a conversation

Good Tape
Secure and automatic transcription

Scribe
ElevenLabs’ SOTA speech-to-text model

InfiniteTalk AI
Next-level conversational voice generation.

MP3 to Text
A lightweight, browser-based MP3 to text converter with TXT and SRT export. Upload an MP3, transcribe, then copy text, download TXT, or export SRT. You can start free with no credit card.

Octave TTS
Generate AI voices with emotional delivery

Y2Doc
Transform YouTube videos into structured documents

Lightricks
This AI platform offers an audio-to-video feature that lets you upload music, voice or sound effects to build a video.

AI Took My Voice. I Want it Back

Beatoven AI
AI composer for crafting the perfect background music

ScriptTimer
A  free script timing calculator built for journalists, podcasters, and broadcast creators. Paste any script and instantly get your exact speaking time, word count, and per-paragraph timing breakdown — no sign-up required. It works in any language and supports slow, normal, and fast speaking speeds for radio segments, podcast episodes, and video packages.


Maibrain
Preserve the voice and experiences of your loved ones so you can interact with them in the future.

​​BeMusic AI: AI Music Generator
AI music generator built for anyone who wants to create original songs without a studio. Whether you need background music for a video, a full-length track for a podcast, or a singing photo for social media, BeMusic AI turns simple text descriptions into studio-quality music in under 30 seconds.

Veracity
A web app built for journalists that transforms interview recordings into a verifiable workspace, where every AI-generated sentence is traceable to the exact moment in the source audio. It combines transcription, summaries and quote extraction with a unique verification layer that lets users click any claim and hear the original source—or see it flagged if it can’t be confirmed. By checking both AI outputs and a journalist’s final article against source material, Veracity ensures accuracy, reduces errors, and protects credibility before publication.

Lyria 3
Google’s new AI music generation model in Gemini

Microsoft VibeVoice
Produces long-form, multi-speaker conversational audio, such as podcasts, from text. It addresses significant challenges in traditional Text-to-Speech (TTS) systems, particularly in scalability, speaker consistency, and natural turn-taking.

AI Translate Video
When you have local MP4/MP3 files that need translation Got a public video URL? Paste it—no download needed. Need separate subtitle files? We’ve got you covered. Want to translate your YouTube videos into multiple languages? Reach global audiences. Keep the original speaker’s voice? Voice cloning can help. Creating multilingual TikTok/Shorts videos to test different markets? Quick and easy.

WhisperTranscribe
Creates podcasts, blogs, and social content

Talktastic
Cleans up audio you record and summarizes it. Website app.

Udio
AI-music generator

Letterly
Captures your voice and lets AI tdo the writing. Phone app and desktop tool. Then tell it how to organize thoughts: Outline/summary for a class or a presentation.

Whisper Web
Free, browser-based transcription tool that converts audio and video files to text in 100+ languages with no signup required. It runs OpenAI’s Whisper model directly in your browser, so interview recordings never leave your device — a meaningful privacy protection for journalists. It’s a practical zero-cost option for reporters transcribing interviews on deadline.

Song Maker
Create complete music with AI. Use our AI song maker and AI song generator to make a song, make song ideas from lyrics to song.

StudioEditor
A free, browser-based audio workspace in which recording, timeline editing, exporting, leveling and automatically ducking music under speech cost nothing. Designed for podcasters, creators, educators, interviewers, journalists and media, it combines a traditional audio editor interface with optional but unbeatable AI-tools, all in your browser.

AI Deepfake Detector
Helps journalists screen video, audio, images, and text for possible AI generation or manipulation before publication. It combines C2PA provenance data, file metadata, and AI analysis to return a risk level with uncertainty notes rather than a definitive verdict; no signup is required, and each visitor receives free checks. DeepFakeCheck does not retain uploads after the request completes, while media sent to Google Gemini is subject to Google’s privacy policy. Two free checks, then a $6.99 monthly fee.

RJI Momentum Podcast
A series of conversations with news industry thought leaders about strategies that can help newsrooms navigate the ever-changing world of technology and audience behavior. In the first season of three episodes available now, RJI Executive Director Randy Picht gets three different newsroom perspectives about generative AI.

FlashEdit
A browser-based AI creative platform for generating and editing videos, images, audio and visual content. Its AI video generator lets users create short-form videos from prompts, images, videos, and other media references using multiple AI video models in one workflow. It is useful for journalists, educators, creators, and content teams who need quick social videos, explainers, story visuals, or media experiments without a complex editing setup.JOURNALISM BASICS

EconoFact
A non-partisan publication by Tufts University Fletcher School, bringing facts and incisive analysis to the national debate on economic and social policies.

Conversent
Turns your studio’s audio and video into knowledge that you control.

scryp
An AI transcription service from Vienna that turns interview, meeting and podcast recordings into accurate text with speaker labels. Files are encrypted in the browser before upload (AES-256-GCM) and transcription runs entirely in scryp’s own data centre in Austria, which makes it a GDPR-safe choice for source protection. It supports 134 languages, subtitle export (SRT/WebVTT) and AI summaries, with a 14-day free trial.

Voice to Notes
A powerful AI-powered tool that instantly converts your voice into clear, structured notes. It offers real-time transcription, automatic summaries, grammar correction, and smart formatting—all accessible across web, Android, and iOS. Designed for students, professionals, and creators, VoiceToNotes.ai helps you capture and organize your ideas effortlessly.

RJI Momentum Podcast
A series of conversations with news industry thought leaders about strategies that can help newsrooms navigate the ever-changing world of technology and audience behavior. In the first season of three episodes available now, RJI Executive Director Randy Picht gets three different newsroom perspectives about generative AI.

LegiTalk from CT Mirror
Tool that summarizes Connecticut legislative meetings and provides transcriptions for reporters to use. They can then click forward to the section of the video to confirm the transcript.

LyricsGenerator.io
Free AI song and lyrics creator

Sonix
Good speech to text tool for transcribing audio.

ChatGPT Advanced Voice
Now available on desktop apps for both Mac and Windows

PDF2MP3

A professional PDF to MP3 converter that turns any PDF into high-quality audio using advanced AI text-to-speech. You can convert PDFs to MP3 in seconds with 61 natural-sounding voices across 8 languages.

Command A Translate
Translation model

Speechmatics
Build voice-enabled apps with Speechmatics’ Startup Program and get up to $50k in credits to deploy to production*

ChatGPT Translate
Translate text, voice, or images across 50+ languages

AssemblyAI
Offers speech recognition and audio intelligence APIs for transcription, speaker diarization, and content moderation.

Nieman Lab: Local Newsrooms Are Using AI to Listen in on Public Meetings

ElevenMusic
Platform for AI song generation, remixing, creator payouts.

LyricsGenerator.io
AI music creation platform, combining a versatile Lyrics Generator with professional Song Production tools.

TranslateGemma
Google’s new family of open translation models

Attesta
An iOS recorder that announces out loud that it is recording, never trains on your audio, and leaves the file yours. The output is built to be usable, not just archived: top-tier models, speaker labels, and the action items, decisions, open questions and key people pulled out accurately. You can extend a recording with Continue, and refine or restyle the summary to match how you write.

Snappixify.net
Convert MP4/MP3, download Instagram videos, and more with our fast online tools. Supports H.264, 4K resolution, and all major formats.

Claude Cowork
Bring Claude Code’s agentic capabilities to everyday tasks

Livekit
New open-source turn detection model for natural voice AI.

Amazon Polly
Deploy high-quality, natural-sounding human voices in dozens of languages

Remento
AI biographer with Speech-to-Story technology that turns recorded interviews into a hardcover book of polished stories

Crush

Autopod
Save hours in your weekly production time in Adobe Premiere Pro with this pack of plug-ins. Free trial to get you started.

Amical
AI voice-to-text app for dictation, meetings and note-taking

GPT-4o-tts and Transcribe
OpenAI’s text-to-speech and speech-to-text tool

Restream.io
One livestream going to more than 30 audiences.

Wordcab One
Transcription, speech intelligence, and summarization, all in one tool.

Wondershare Virbo
A multiple languages video translator that can generate AI voice clones and lip syncs. It automatically generates subtitles, avatars scripts with AI.

MiniMax T2A-01 HD
Text-to-audio model enabling voice cloning with just 10 seconds of audio

AI Convert Hub
A fast, secure online file conversion and compression platform for video, audio, and images, supporting more than 200 formats with no software installation. It offers batch uploads, AI-optimized compression to preserve quality and advanced codec-aware controls for power users. With a free tier up to 1GB, multi-language support, and SEO-friendly positioning, it helps creators and teams quickly convert or compress files entirely in the browser.

AI Song Generator Free
Easy-to-use platform that creates high-quality, royalty-free music in various styles, perfect for videos, ads, and creative projects.

Notta Showcase
Translate videos into 15-plus languages while retaining the original voice.

InfiniteTalk AI
Create infinite‑length talking videos from any video or image. InfiniteTalk AI delivers razor‑accurate lip sync, expressive full‑body motion, and rock‑solid identity preservation—powered by next‑gen sparse‑frame technology. 

Mumble Note
AI-powered voice notetaker designed for journalists and creators to instantly transform interviews, meetings, or idea sessions into clean, structured notes and action items. It automatically extracts key points, generates summaries, and tags content—so you can stay on story instead of scrambling for notes. With encrypted processing and support for images, links, and voice-to-text in 40+ languages, it blends privacy, versatility, and speed for every reporting workflow.

Tomedes AI Transcription
Supporting formats like MP3, MP4, WAV, and nearly 100 languages, it’s perfect for transcribing interviews, meetings, and lectures.

Stable Audio Open Small
Text-to-audio model for music sample

Nova-3 by Deepgram
New voice AI model for real-time multilingual transcriptions in real-world enterprise use cases

CloneDub
Convert audio into other languages using the same voices. Only audio files, YouTube, or audio links less than 15 minutes will work. Note: Many ethical concerns with this tool. Fact-check any translations for accuracy.

Poynter: AI Detection Tools for Audio Deepfakes Fall Short
Experts test four tools and show options on what to do.

Conversational AI
ElevenLabs tool that allows users to seamlessly add voice capabilities in 31 languages 

Clipto
Convert audio to text in 99 languages.

TalkText
An AI-powered dictation tool that transforms speech into text

1B CSM
Text and audio-to-speech model

SynthID
Audio tool from Google Deepmind

Resemble.ai
Deepfake audio detection

Transkriptor
Convert audio to text quickly

Pindrop
Realtime audio deepfake detection

Tad.ai
Create original music from prompts

Voiser AI
Transcribe, summarize, and translate videos and recordings

Hume AI Voice Control
Helps developers create consistent, custom AI voices by adjusting 10 sliders and settings.

AgentPlace
Create AI-driven websites and apps through simple text instructions

DiffRhythm
Generate complete 4-min songs w/ vocals in just 10 seconds

Akool.com
Create video, audio and characters. Also translate audio into other languages

Adobe Enhance Speech
Upload any audio recording with background noise and immediately get a clean version.

Video to Text
Video to text is an ai-powered transcription service that converts video and audio files into clean, exportable text. the product is designed for creators, teams, and individuals who need fast, accurate speech-to-text conversion without setting up their own transcription pipeline.

Read PDF Aloud
Let AI Read Your PDF Aloud with Natural Voice. Free AI-powered PDF speaker. Listen to any PDF document with natural-sounding voices in 140+ languages.Read your PDF aloud with one click.

Good Tape
A security-first AI-powered transcription tool created by journalists and known for its fast turnaround and unprecedented accuracy, when it comes to audio and video transcriptions. Trusted by over 2.5 million users worldwide, GoodTape has transcribed more than 3 million hours of audio and saved 12 million hours of work globally. Good Tape understands the importance of confidentiality when working with sensitive sources and materials and does not use the customer’s transcription files for AI learning of any kind.

ChatGPT Record
Capture, summarize, and transcribe audio with ChatGPT

Magenta RealTime
Google’s new open-weighted live music model

Sonix
Powerful transcription tool. 30-minute free limit and support more than 40 languages.

 

Fireflies
Its free plan offers unlimited transcription, with storage limits.

All Voice Lab
Delivers AI-powered voice cloning capabilities

ACE Studio
AI workstation to generate studio-quality singing vocals

Projects by ElevenLabs
Build  long-form audio 

Stock Music GPT
Instant royalty-free stock music, sound effects and song covers, generated by AI.

Workflowy
Powerful outliner helps to jot down ideas, thoughts, writing topics and block drafts and to easily re-arrange and organize them. Can also be used to manage tasks and projects. Intuitive. Fast. Definitely my favorite outliner. Free version (all features with monthly limit). Pro version $8.99 a month.

MusicFX
Google’s free text-to-song creation tool

Orpheus TTS
Open-source text-to-speech AI with natural emotion


Talo
Real-time voice translator for video calls

GitPodcast
Build podcasts to understand GitHub repositories

Newsroom Robots Podcast: Congresswoman Anna Eshoo on Shaping the Future of AI Policy



Peech
Effortlessly transforms any text into incredibly realistic AI-generated audio. Peech supports over 50 languages, including English, French, German, Italian, Spanish, and more. 

Listnr
A free voice AND video generator in multiple languages

Papercup
AI video and audio dubbing tool

MusicFX
Google tool that lets you describe what you want and it will write a song.

Eddy
Free transcription tool from Headliner app.

Speech-02
Minimax’s text-to-speech AI supporting over 30 languages

Vapi
Build voice AI agents

Everlit Audio
Audio creation tool. It has a WordPress plug-in, embeddable player, APIs, voice replication and more.

Eleven v3
SOTA text-to-speech model with support for 70+ languages

Bland TTS
Voice AI with enhanced control

HunyuanVideo-Avatar
Multi-character talking videos from audio

Musick AI
Creating high-quality, original music across various genres. user-friendly interface, time-saving capabilities, AI filters, mood templates, and interactive delivery.

AudioRead
Use AI to listen to articles, PDFs, emails, etc. in your podcast player. Read while walking, driving, cleaning and more.”Read” while walking, driving, cleaning, and more

Superwhisper
Fast and accurate voice to text

Udio Audio Inpainting
Select a portion of an AI-generated music track and regenerate it. Be careful with music rights using this tool. 

Stability AI’s Stable Audio Open
Generates up to 47-second audio samples based on text descriptions. It’s trained on thousands of  royalty-free music samples.

Transcript LOL
Transcribes podcasts, videos and meetings

GitPodcast
Generate engaging podcasts to understand GitHub repositories

Supertone Shift
A real-time voice changer

ElevenLabs Audio Native
Add narration to your blog or news site

Adobe Express Animate from Audio
Upload any audio and animate characters with AI

AI Cover Art
A streamlined platform that uses artificial intelligence to create high-quality music covers quickly. The tool transforms songs with AI-generated vocals that sound natural and professional. Users can select from different voice styles, adjust vocal parameters, and export their creations in minutes. Perfect for musicians, content creators, and music enthusiasts who want to explore new versions of their favorite songs without complex equipment or technical skills. 

YouTube Transcription Generator
A free, no‑login tool for instantly generating transcripts from any YouTube video. Users can extract subtitles in multiple languages, search within the transcript, and copy or download TXT, SRT, or VTT files, with or without timestamps. Designed for creators, students, researchers, and editors, it streamlines workflows by making it fast and distraction‑free to grab accurate transcripts and jump to exact moments in a video.

Tomedes AI Transcription
A free, browser-based speech-to-text tool that converts audio files into text using multiple leading AI engines—Whisper, Gemini, and Amazon Transcribe. It generates up to three transcription results per file, allowing users to compare and select the most accurate version. With no sign-up, no ads, and multilingual support, it’s ideal for journalists, podcasters, and researchers.

SeedanceHub | AI Video Generator
Transforms text prompts or images into high-quality videos. It offers smooth motion, vivid visuals, and fast generation speeds.

Chuchotis
A whispered-based application on MacOS (Apple silicon) for 100% confidential voice-to-speech transcription, developed by a french journalist, and heavily tested by his colleagues. It offers many tools to improve transcription (prompt, hallucinations & incomplete segment, batch processing, tools for edition… Localized in English, french and spanish, works on 100 languages.

EVI 3
Create custom voices through speech-to-speech interactions

Conversational AI 2.0
ElevenLab’s updated AI voice agent platform

PlayDiffusion
PlayAI’s open-source audio inpainting model

Cassette AI
Music-generation tool.

Pdf2audio
Open-source tool that converts PDFs to podcasts, lectures, summaries and more

Minutes
Automate meeting notes and transcriptions with AI

Alphana
Convert podcasts and video into several short, viral pieces.

Eleven Labs: Multispeaker Podcasts
Works like Notebook LM

Muzix
Transform text into music with AI music generator. Create custom songs and instrumental tracks in minutes, no musical experience needed.

Plus AI
Google Slides tool generates presentations from any text input.

SOUNDRAW
Uses rights-free audio to build music.

OpenMusic
Create custom tunes from text descriptions

Udio Lyric Editor
Create and refine song lyrics based on melody

Soundraw
AI music generator

Mubert
Generative AI music for videos and podcasts

Artist.io
Create voiceovers for video using AI

VoiceType
Most professionals spend 5-10 hours a week typing. But this text-to-voice tool speeds that up.

Songscription
Turning audio clips instantly into sheet music

Suno
Generate a song from any sound

Voice Isolator by Eleven Labs
Strip background noise from interviews, films and podcasts

SpeakHints
Real-time AI speech copilot for any situation

Minvo
AI video shorts platform for podcasts and more.

Wondercraft AI Audio
Creates hyper-realistic audio ads

Soundraw
Generate music/beats using AI.

Jotify
Turn research into stories and audio

Mixhalo
Audio translation on your phone

ElevenLabs Text to Sound Effects
Generate any sound from a prompt

Pika Labs
Add sound effects to your videos

Synthflow AI
Create conversational AI voice agents without coding.

Auphonic
AI sound engineer for podcasts, etc.

MacWhisper Pro
Free transcription tool with strong privacy

Jellypod
Transform your inbox into a personalized daily podcast 

izTalk
Translation and speech recognition

OpenVoice
A free tool to clone voices 

Letterly
Record your voice and let AI turn it into well-written text

LALAL AI
Remove voices and instrumental audio splitter

Deepgram Aura
A text-to-speech API built for real-time conversations.

CleanVoice
Cleanvoice is an artificial intelligence which removes filler sounds, stuttering and mouth sounds from your podcast or audio recording

Coval
Make voice and chat agents with seamless simulation

PopPop AI Vocal Remover
Free AI-powered tool to separate or remove vocals from any song

ArticleReader
Turns text into captivating audio

EVI-2 by Hume
AI speech with emotion, using several voices and personalities

CloneDub
Automatically dubs your videos and podcasts

Dubformer
AI-powered dubbing solution with broadcast quality 

Video Insights
Video and audio summarization and transcription

Dubformer
AI audio dubbing tool

MixAudio
AI music generator and editor

Poynter: How Text-to-Speech Technology Can Help Journalists Avoid Copy Errors

Udio New Features
Generate AI music longer than 2 minutes and extend tracks up to 15 minutes

Hive AI Detection
Check photos, video, audio and text. Freemium tool and you can request a demo.

PocketPod
Creates personalized podcasts that match your interests.

Replica Voice AI
Ethical voice AI for creators and business

Chuchotis
Chuchotis (French word for whisper) is a fee-based speech-to-text application based on OpenAI’s whisper. It runs fully locally on any M-silicon-based Mac, to preserve privacy of recordings and transcriptions.

FlowTunes
Endless AI-generated music for focused work

SpeechTexter
Type with your voice

Beatoven AI
Create royalty-free background music.

SpeakAI
Get transcription, research, data analysis and NLP software.

Udio Audio Prompting
Upload a sound and let AI generate a song

AI Jukebox
In-browser text-to-music generation

Soundry AI
AI sound sample VST for music creation and DJing

AIornot.com

Ramble Fix
An audio to text tool. Speak your mind and it will paraphrase/rewrite what you say. Five free uses per month, then $10 monthly for more recordings.

Teachable Machine
Teachable Machine is a web-based Google tool that makes creating machine learning models fast, easy, and accessible to everyone. Train a computer to recognize your own images, sounds, and poses. A fast, easy way to create machine learning models for your sites, apps, and more – no expertise or coding required.

Adobe Speech Enhancer
Speech enhancement makes voice recordings sound as if they were recorded in a professional studio. Be careful with ethics and edits with this tool.

Google AudiopaLM
A large language model for speech understanding and generation. AudioPaLM fuses text-based and speech-based language models, PaLM-2 and AudioLM [Borsos into a unified multimodal architecture that can process and generate text and speech with applications including speech recognition and speech-to-speech translation.

Easy_Peasy.ai
Create images with text, transcribe audio and create audio with this Swiss Army knife of an AI tool

Rizzle AI
Convert text and podcasts into captivating videos

Toasty AI
AI content creation for podcasts

CastMagic
Turn audio into content

Audio Notes AI
Organize thoughts into structured notes with AI

Article Audio
Convert your article to audio. Free model with paid upgrade.

Pickle
Lifelike AI clones lip-syncing to your voice in real-time. Be careful with ethics if you are using this tool in a journalistic way.

Wordtune
Rewrite your thoughts in different styles.

Castmagic
Turn long-form audio into ready-to-use content assets, instantly. Requires a log-in to PartnerSnack first.

Co-Producer Pack Generator
Generate authentic music sample packs

Music Maestro GPT
A ChatGPT GPT assistant for music creation

MusicFX DJ
Google has added ‘DJ Mode’ in MusicFX, the generative text-to-music tool powered by Google’s MusicLM

PDFToMP3
Transforms PDFs into MP3 formats

SwellAI
AI assistant for podcast producers

Nendo
Open-sourced AI music generation models

Musicgen-remixer
Remix music into other styles

Respeecher
Hollywood-grade AI voice generation

OpenAI Advanced Voice
Creat conversations with AI, available forChatGPT Plus subscribers

Whisper WebGPU
A real-time in-browser speech recognition with OpenAI Whisper. The model runs fully on-device and supports  transcription across 100 different languages.

MAGNeT
A high-quality text to sound and text-to-music

Podsqueeze 2.0
Podcast content repurposing

Hypernatural
Use it to generate audiogramshttps://www.podmind.ai/

Dub AI
AI-powered voice cloning and translation

Read This AI
Transform text into high-quality audio effortlessly

Voicenotes
Notetaker that transcribes and provides info on your thoughts

izTalk
Overcome language barriers with instant AI translation 

Text Reader AI
Convert text to speech with free, realistic AI voices

BlueDotHQ
AI meeting recorder backed by Google

AudioPen
Just hit record. Then start talking. AudioPen will clean things up when you’re done.

TalkNotes
Transcribes your voice notes in more than 50 languages

Lalal
Extract vocal, accompaniment and various instruments from any audio and video. One-time fees start at $15

Zealous
Convert audio files into social media posts. Tool is marketing-driven but has social media desk benefits. Free with reasonable upgrades.

Voicemod
An AI voice creator that lets users make their voices by changing gender, age and tone.

Mubert
Create soundtracks for your projects with AI. Freemium account

Sumly
Audio and podcast summary tool

Director Mode by Wondercraft
Fine-tune and direct AI voices through prompts

Audeus
Text-to-speech AI for PDFs, email, etc.

Voicify
Create AI covers using AI in seconds with Voicify, with hundreds of community-uploaded AI voice models available for creative use now.

Krisp
Noise-canceling app

Altered
Voice-changing software. Be sure to use this ethically and carefully.

VoiceMod
Free real-time voice-changer. Be sure to use this ethically and carefully.

BeyondWords
Text-to-speech tool. Free account with paid upgrades.

Samplab
Edit audio with AI 

Vocal Remover
Separate voice from music out of a song free with powerful AI algorithms

AI Jingle Maker
Create jingles, radio sweepers, podcast intros, audio promos and more.

AssemblyAI | Playground
Production-ready AI models for speech recognition, speaker detection, audio summarization, and more through our API. Quickly test below using any YouTube link, audio file or video file.

RadioInfo Australia: Embracing the AI Wave: How Media Companies Can Successfully Integrate AI Technologies

Narakeet
Text-to-speech tool for creating voiceovers. Pay by the length of the audio.

Cassette.ai
Music generation and text-to-audio tool

CastMagic
Upload and .mp3 and download a transcript.

Wavel
AI video dubbing with realistic voices, subtitles, accents, etc.. 

Flowjin
AI-generated short video clips for audio or video podcasts

Songburst
Mobile AI music generation app with prompt enhancer

Byrdhouse AI 2.0
Multilingual video call AI interpreter

Listnr
Free AI voice and video generator, choose from 900+ voices in 142 languages. Get started for free, download in MP4/MP3/WAV formats.

Listnr
Text-to-voice and text-to-video tool.

Assembly
Exposes AI models for speech recognition, speaker detection, speech summarization, and more.

Khanmigo
AI teaching assistant from Khan Academy

Poised
AI-powered communication coach that helps you speak with confidence and clarity. Private and secure, an essential tool for digital-first workplaces. Free trial with personalized plans.

Kits.ai
Create AI voices or modify your own with a library of commercial use and officially licensed artist voices.

Brain.fm

BeyondWords
Convert writing into audio. Free for up to 30,000 words, then paid.

Notta
Notta is an AI-based voice-to-text transcription software that supports 104 languages. Notta excels at transcribing and summarizing audio or video files, online meetings, and voice recordings. Additionally, Notta offers a suite of team alignment features, enabling users to schedule and transcribe Zoom, Google Meet, and Teams meetings, among other functionalities.

Wondercraft AI

Boomy – Make Generative Music with Artificial Intelligence

Whisper
Popular AI tool for transcription.

Fathom
Note-taking and transcription tool for Zoom

Media.io
A collection of AI tools ranging from video/image enhancement to audio.

Wellsaid
Creates texts from voiceovers in seconds

Transistor
Quick and precise speech-to-text conversion for podcasts

PolyAI Pheme
Generate conversational voices for phone-call apps

Nolan Free Script-Writing Software

Decoherence
Make AI music videos

Castmagic
Create podcast notes by uploading an .mp3 file download all your post-production content.

Kaiber
An AI creative lab on a mission to unlock creativity through powerful and intuitive generative audio and video. Upload a song, add a touch of your artistic style, and let its audio analysis technology do the visual work. Monthly plans ranging from $5 to $25 with a free trial.

GIJN: Podcasts for Investigative Journalists
Global podcasts in a variety of languages

Suno.ai
Create songs 

AI Phone
Provides live transcription and teal-time translation during phone calls using AI.

TurboScribe
Converts all audio and video files into transcripts

Songburst AI
AI-driven music generator

Papercup
AI dubbing for video

Respeecher
Clones voices for content creators

Bloks
Free desktop note-taking app

Wavel.ai
Dub videos in more than 30 languages. Free with paid upgrades. Be careful with ethics with this tool.


AI PODCASTING TOOLS

Adobe Podcast
AI audio recording and editing, all on the web.

Podcastle
A cloud recorder and AI-powered editor that lets you record a remote interview, edit and mix

Podcraftr
Repurpose blog posts, email newsletters, training materials into podcasts.

Wondercraft
Streamline your podcast with only text.

Auphonic
Upload your file, and Auphonic will automatically optimize your recording and clean it.

Castup
A professional audio and video editing service for podcasters and podcast networks. Castup offers subscriptions starting at $30 per episode, or 40 cents per published minute. Castup also offers a ChatGPT-powered podcast assistant called Castup AI, which can help you record and promote episodes.

Streamyard
A professional live streaming and recording studio in your browser. Record your content, or stream live to Facebook, YouTube, and other platforms. Freemium account with paid models starting at $20 a month.

Audioread: Read. In Audio.
Turns articles, PDFs, etc. into podcasts using this AI-driven tool

Cleanvoice.ai
 AI audio tool that can remove unwanted sounds and speech imperfections.

Listener.fm
Use AI to generate podcast titles, descriptions, and show notes in seconds

Promptcast
Summarizes any podcast with AI

Listnr
Paste text into the text-to-speech converter, and the app will convert it into one of their 600 voices.

Alitu
AI-powered one-stop shop for podcast production. Record remotely and edit.

Podwise
A learning app for podcast listeners

Newsroom Robots: Podcast About AI

Google’s Music LM
Create your own music in Google’s test kitchen.

Mubert
AI-generated music

The Essential AI Toolkit for Journalists and Content Creators
All the tools listed are free to use and open-source. Audio, images, web scraping and other tools

Boomy
AI music generator. Free version with paid models ranging $10 to $30 a month.

Listen Notes
A ChatGPT plug-in that allows you to search for podcasts by person or topic.