What happened: Mitsubishi Electric and its U.S. subsidiary Mitsubishi Electric Research Laboratories have developed a single AI model that can separate and extract specified sounds from mixed audio based on a text prompt, handling tasks like separating multiple speakers, speech enhancement, and extracting environmental sounds.


EDM producer Max "H4RRIS" Harris and Italian turntablist-turned-producer Nihil Young are publicly calling out…

Beatport now bans tracks made entirely or mostly by AI

The Fire and Disaster Management Agency plans to launch a model project in fiscal 2027 to use AI in handling 1…
AWS announced an integration where Amazon Quick, an agentic AI workspace, connects to fal's generative media p…

Leafnet and BBIX began collaborating in August 2026 to build a new service that combines voice and AI

Deepgram has introduced two capabilities for its speech-to-text and text-to-speech models on Amazon SageMaker…

AWS, NVIDIA, and Heidi Health published a solution that cuts GPU infrastructure for automatic speech recogniti…

Plaud introduced the Plaud One Explorer Edition, AI earbuds that record, transcribe, and summarize conversatio…

Plaud, which has over 2.5 million users, has launched Plaud One, earbuds that record calls and use a case to r…

Google has launched Gemini 3.5 Transcribe, a speech-to-text model for real-time transcription that auto-correc…

Relay, founded by two former Nothing employees, is debuting a dedicated microphone for high-fidelity voice-to-…

Netflix released the trailer for 'Wonka's The Golden Ticket,' a nine-episode reality competition show streamin…

Netflix is making a competition reality show based on Willy Wonka & the Chocolate Factory, and it will use AI…

Google announced Gemini 3.5 Transcribe, an AI model that edits out “ums” and corrections to output polished te…

Particle, the AI newsreader startup founded by former Twitter engineers, introduced Radar, a podcast search en…

Stability AI has raised $76 million from a group of investors that includes Sony Music Group, Universal Music…
Australia’s recorded music industry will ban tracks wholly generated by AI from official charts starting next…

AWS published a walkthrough for building a voice ordering system that answers a restaurant's phone line, greet…

Apple Music will add “Made With AI” labels to songs where “a material portion” was created using AI, starting…

A software engineer created Schmaudio, a platform that generates interactive audio stories where listeners mak…

Researchers tested 11 widely used open-source speech recognition models and found that several of the highest-…

Apple researchers applied an iterative pseudo-labeling training approach to Mandarin-English code-switching AS…

Adobe is releasing three AI audio tools — Generate Music (royalty-free music for videos), Generate Speech (scr…

Adobe announced general availability of audio generation capabilities in Firefly, its creative AI suite
Scammers are increasingly using AI-powered voice-cloning and deepfake technology to imitate trusted contacts…

A Reddit discussion raises concerns about AI voice cloning targeting diplomats and government officials, notin…

Wispr AI, which makes Flow—a dictation tool that converts spoken words into clean text—raised $280 million in…
Enterprise voice AI company Regal Voice has integrated its autonomous AI voice agents into Five9's contact cen…
VocalCode, a desktop dictation tool, lets developers speak code and prose directly into editors, terminals, an…

Jakob Jordan, a young man with autism and apraxia who cannot reliably speak aloud, starred in an opera called…

A developer released Dictata v0.1.0, a Windows application that transcribes speech locally using Whisper (an A…

Suno rolled out Studio 2.0 with a beta chat feature that lets Premier subscribers talk to the platform like a…

Suno released Studio 2.0, a major upgrade to its generative AI music tool that adds MIDI support (its most req…

Twitch has added a toggle in its Security and Privacy settings that lets streamers opt out of having their con…

Sandbar, maker of the private voice ring Stream, raised $36 million total to date, including a $23 million Ser…

Sandbar, maker of the voice-enabled ring Stream, raised $36 million to date, including a $23 million Series A…

Guitar company D'Addario has admitted that Suno AI was used to regenerate a track in a promotional video for n…

Five minutes each morning. That's the whole AI news habit.
The day's essentials from 200+ sources, delivered to Email, LINE, or Slack. Free, always.
30,000+ monthly readersAfter former lead writer Stella Sacco posted on Bluesky that Saber replaced her with ChatGPT midway through de…

Meetily, a free open-source meeting assistant, lets users record and transcribe video calls directly on their…

LA rapper Fenix Flexin has reversed course and now claims he "never said I didn't use AI" to make his hit song…

AI music generator Suno announced new policies in response to legal pressure and spam abuse, including stricte…

Musicians including MattstaGraham and River / iamriverhawk are uploading deliberately strange or poorly-made s…

ChatGPT's standard voice mode no longer operates in turn-based format, allowing it to be interrupted mid-respo…

Suno is implementing watermarks on AI-generated music and tightening rules against misuse, including limits on…

Suno announced new watermarking and fingerprinting technology, along with updated download policies, to limit…

A developer has built LiveTranscriber, an open-source iOS app that runs speech and language models entirely on…

Wispr Flow, a voice dictation startup, launched Notetaker, a feature that transcribes meetings in real time us…

OpenAI president Greg Brockman said at a media roundtable on July 23 that voice will be the interface of the f…

Wrinkles, available on iOS and Android, uses location data to automatically surface audio stories about places…

Spotify announced during its Q2 earnings call that Merlin, a licensing partner for independent labels and dist…

OpenAI has introduced GPT-Live, a system that enables continuous voice interaction with AI using a turnless sp…

AGIBOT's WITA-Omni Preview foundation model achieved an average accuracy of 85.21 percent on the Daily-Omni au…

Smallest.ai, founded in late 2024, raised $13 million in a Series A round led by Seligman Ventures, with parti…

Friend, the AI pendant startup, relaunched its product with a built-in speaker that can talk to users, raising…

Fish Audio, an AI startup founded by former Nvidia researcher Shijia Liao, raised $52 million in seed funding…
Google released Lyria 3.5, a music generation model that produces more natural-sounding melodies, better lyric…

OpenAI released GPT Transcribe and GPT Live Transcribe, two speech recognition models via its API

OpenAI announced GPT Transcribe, a speech-to-text model that processes completed audio files, streamed file tr…

Palo Alto-based Fish Audio, which has over 8 million users and generates $21 million in annual recurring reven…

An expert committee of Japan's Justice Ministry approved a draft report Monday establishing that unauthorized…

Listen Notes, a podcast search database, reports that over 30% of newly created podcasts consist of AI-generat…

Deepgram has integrated IAM temporary delegation, a new AWS IAM capability, into its support workflow for cust…

Alex Smola, former distinguished scientist at Amazon, founded Boson AI to release Higgs RealTime, a speech-to-…

A developer has released ARIA, a voice-controlled security operations cockpit (SOC) that runs on local hardwar…

OpenAI updated its ChatGPT desktop app on Thursday to support ChatGPT Voice, allowing users to speak commands…

Anthropic expanded Claude's voice mode to its Opus and Sonnet models (previously limited to Haiku), with the a…

OpenAI has released ChatGPT Voice on desktop and laptop computers, extending the GPT Live voice technology it…

Skipper is an AI companion designed to work without internet, offering voice interaction, custom memory, HD vi…

Anthropic has made its Opus and Sonnet models available in voice mode—until now limited to the lighter Haiku m…

Anthropic updated Claude's voice mode to let users choose between Opus, Sonnet, and Haiku models, with the mod…

German AI company Black Forest Labs released Flux 3, a multimodal foundation model that learns from images, vi…

Get curated AI news from 200+ sources delivered daily to your inbox. Free to use.
30,000+ monthly readers
Get Started FreeFree · takes 30 seconds · unsubscribe anytime