AIToday

Catch up on AI in one minute, every morning.

Start free30,000+ monthly readers

Updated: 10h ago

Today's Pick

Mitsubishi Electric develops task-general sound separation AI

What happened: Mitsubishi Electric and its U.S. subsidiary Mitsubishi Electric Research Laboratories have developed a single AI model that can separate and extract specified sounds from mixed audio based on a text prompt, handling tasks like separating multiple speakers, speech enhancement, and extracting environmental sounds.

Top Companies AI10h ago
Mitsubishi Electric develops task-general sound separation AI

Latest AI News

Earlier

AWS Quick and fal unite for agentic creative workflows

AWS announced an integration where Amazon Quick, an agentic AI workspace, connects to fal's generative media p…

Amazon AI Blog4d ago

Leafnet and BBIX team up on voice-AI services

Leafnet and BBIX began collaborating in August 2026 to build a new service that combines voice and AI

Top Companies AI4d ago

Deepgram expands SageMaker AI observability metrics

Deepgram has introduced two capabilities for its speech-to-text and text-to-speech models on Amazon SageMaker…

Amazon AI Blog4d agoAlso reported by Top Companies AI

NVIDIA MPS slashes ASR GPU needs by 75% on AWS

AWS, NVIDIA, and Heidi Health published a solution that cuts GPU infrastructure for automatic speech recogniti…

Amazon AI Blog4d ago

Plaud launches AI earbuds for conversation recording

Plaud introduced the Plaud One Explorer Edition, AI earbuds that record, transcribe, and summarize conversatio…

The Verge AI4d ago

Plaud launches AI note-taking earbuds with eSIM case

Plaud, which has over 2.5 million users, has launched Plaud One, earbuds that record calls and use a case to r…

TechCrunch AI4d ago

Google launches Gemini 3.5 Transcribe in 85 languages

Google has launched Gemini 3.5 Transcribe, a speech-to-text model for real-time transcription that auto-correc…

THE DECODER4d ago

Startup Relay builds dedicated mic for voice-to-text

Relay, founded by two former Nothing employees, is debuting a dedicated microphone for high-fidelity voice-to-…

WIRED AI4d ago

Netflix's 'Wonka's The Golden Ticket' Trailer Uses AI to Recreate Gene Wilder's Voice

Netflix released the trailer for 'Wonka's The Golden Ticket,' a nine-episode reality competition show streamin…

Top Companies AI5d ago

Netflix Uses AI to Recreate Gene Wilder's Voice for Wonka Show

Netflix is making a competition reality show based on Willy Wonka & the Chocolate Factory, and it will use AI…

Top Companies AI5d ago

Google unveils Gemini 3.5 Transcribe for cleaner voice input

Google announced Gemini 3.5 Transcribe, an AI model that edits out “ums” and corrections to output polished te…

Ars Technica AI5d ago

Particle launches Radar podcast search engine

Particle, the AI newsreader startup founded by former Twitter engineers, introduced Radar, a podcast search en…

TechCrunch AI5d ago

AMD and top record labels back Stability AI's $76M round

Stability AI has raised $76 million from a group of investors that includes Sony Music Group, Universal Music…

SiliconANGLE AI6d ago

Australia bans AI-generated music from charts

Australia’s recorded music industry will ban tracks wholly generated by AI from official charts starting next…

Fortune AI6d ago

AI host takes restaurant phone orders via Amazon Connect

AWS published a walkthrough for building a voice ordering system that answers a restaurant's phone line, greet…

Amazon AI BlogAug 24, 2026

Apple Music to Label AI-Made Tracks Later This Year

Apple Music will add “Made With AI” labels to songs where “a material portion” was created using AI, starting…

Top Companies AIAug 23, 2026

Software Engineer Builds Interactive Audio Stories Powered by AI

A software engineer created Schmaudio, a platform that generates interactive audio stories where listeners mak…

Hacker NewsAug 21, 2026

Top speech AI models show signs of benchmark gaming, study finds

Researchers tested 11 widely used open-source speech recognition models and found that several of the highest-…

Hugging Face BlogAug 21, 2026

Apple improves code-switching speech recognition with iterative pseudo-labeling

Apple researchers applied an iterative pseudo-labeling training approach to Mandarin-English code-switching AS…

Apple Machine LearningAug 21, 2026

Adobe Firefly launches AI audio tools, adds Google's Gemini Omni Flash

Adobe is releasing three AI audio tools — Generate Music (royalty-free music for videos), Generate Speech (scr…

THE DECODERAug 20, 2026

Adobe adds music, speech, sound effects to Firefly AI tool

Adobe announced general availability of audio generation capabilities in Firefly, its creative AI suite

SiliconANGLE AIAug 20, 2026

AI voice-cloning fuels surge in social engineering attacks

Scammers are increasingly using AI-powered voice-cloning and deepfake technology to imitate trusted contacts…

Top Companies AIAug 19, 2026

Deepfake voices pose growing risk to diplomatic communications

A Reddit discussion raises concerns about AI voice cloning targeting diplomats and government officials, notin…

r/artificialAug 18, 2026

Wispr raises $280M for AI speech-to-text at $2B valuation

Wispr AI, which makes Flow—a dictation tool that converts spoken words into clean text—raised $280 million in…

SiliconANGLE AIAug 17, 2026

Regal Voice AI agents now available via Five9's contact center platform

Enterprise voice AI company Regal Voice has integrated its autonomous AI voice agents into Five9's contact cen…

SiliconANGLE AIAug 17, 2026

VocalCode brings push-to-talk dictation to code editors, $4.99 one-time buy

VocalCode, a desktop dictation tool, lets developers speak code and prose directly into editors, terminals, an…

Hacker NewsAug 17, 2026

AI gives nonspeaking opera singer expressive voice

Jakob Jordan, a young man with autism and apraxia who cannot reliably speak aloud, starred in an opera called…

Hacker NewsAug 16, 2026

Dictata: Local Whisper voice dictation for Windows, no data leaves machine

A developer released Dictata v0.1.0, a Windows application that transcribes speech locally using Whisper (an A…

Hacker NewsAug 15, 2026

Suno Studio 2.0 adds chat interface, unrestricted export for paid users

Suno rolled out Studio 2.0 with a beta chat feature that lets Premier subscribers talk to the platform like a…

THE DECODERAug 13, 2026

Suno adds MIDI, automation, AI chat to Studio 2.0

Suno released Studio 2.0, a major upgrade to its generative AI music tool that adds MIDI support (its most req…

The Verge AIAug 13, 2026

Twitch streamers can now opt out of Amazon AI training

Twitch has added a toggle in its Security and Privacy settings that lets streamers opt out of having their con…

The Verge AIAug 12, 2026

Sandbar's voice ring Stream scores $23M Series A, bets on human control

Sandbar, maker of the private voice ring Stream, raised $36 million total to date, including a $23 million Ser…

TechCrunch AIAug 12, 2026

Sandbar raises $36M for voice ring Stream, bets on push-to-talk to survive AI hardware graveyard

Sandbar, maker of the voice-enabled ring Stream, raised $36 million to date, including a $23 million Series A…

TechCrunch AIAug 12, 2026

D'Addario admits AI music in promo video after two weeks of denials

Guitar company D'Addario has admitted that Suno AI was used to regenerate a track in a promotional video for n…

The Verge AIAug 12, 2026

Five minutes each morning. That's the whole AI news habit.

The day's essentials from 200+ sources, delivered to Email, LINE, or Slack. Free, always.

30,000+ monthly readers
Get Started Free

Saber denies replacing writers with ChatGPT on Rideshare Stimulator game

After former lead writer Stella Sacco posted on Bluesky that Saber replaced her with ChatGPT midway through de…

The Verge AIAug 12, 2026

Meetily Brings Free Meeting Transcription Without Upload or Subscription

Meetily, a free open-source meeting assistant, lets users record and transcribe video calls directly on their…

WIRED AIAug 9, 2026

LA Rapper Fenix Flexin Now Admits Using AI for Hit Song 'Rubberz'

LA rapper Fenix Flexin has reversed course and now claims he "never said I didn't use AI" to make his hit song…

The Verge AIAug 7, 2026

Suno tightens download rules, transparency tools to address copyright claims

AI music generator Suno announced new policies in response to legal pressure and spam abuse, including stricte…

THE DECODERAug 7, 2026

Artists release intentionally bad music to confuse AI training

Musicians including MattstaGraham and River / iamriverhawk are uploading deliberately strange or poorly-made s…

Hacker NewsAug 7, 2026

ChatGPT Standard Voice Mode Loses Turn-Based Feature

ChatGPT's standard voice mode no longer operates in turn-based format, allowing it to be interrupted mid-respo…

r/artificialAug 7, 2026

Suno adds watermarks to AI music, seeks legal ground amid lawsuits

Suno is implementing watermarks on AI-generated music and tightening rules against misuse, including limits on…

Ars Technica AIAug 6, 2026

Suno plans watermarking and download limits to curb spam AI music

Suno announced new watermarking and fingerprinting technology, along with updated download policies, to limit…

The Verge AIAug 6, 2026

Open-source iOS app runs speech AI models fully offline on iPhone

A developer has built LiveTranscriber, an open-source iOS app that runs speech and language models entirely on…

r/MachineLearningAug 5, 2026

Wispr Adds Meeting Recorder, Joining AI Notetaker Rush

Wispr Flow, a voice dictation startup, launched Notetaker, a feature that transcribes meetings in real time us…

WIRED AIAug 5, 2026

OpenAI bets voice will replace typing; Brockman: clicking is 'a phase'

OpenAI president Greg Brockman said at a media roundtable on July 23 that voice will be the interface of the f…

Fortune AIAug 4, 2026

Wrinkles app turns phone into AI audio tour guide for local stories

Wrinkles, available on iOS and Android, uses location data to automatically surface audio stories about places…

TechCrunch AIAug 4, 2026

Spotify expands AI remix tool with Merlin, bringing 30,000+ independent labels

Spotify announced during its Q2 earnings call that Merlin, a licensing partner for independent labels and dist…

TechCrunch AIAug 4, 2026

OpenAI launches GPT-Live for real-time voice conversations

OpenAI has introduced GPT-Live, a system that enables continuous voice interaction with AI using a turnless sp…

OpenAI BlogAug 3, 2026

AGIBOT's audio-visual model tops Daily-Omni benchmark at 85.21%

AGIBOT's WITA-Omni Preview foundation model achieved an average accuracy of 85.21 percent on the Daily-Omni au…

Robotics & Automation NewsAug 3, 2026

Smallest.ai raises $13M for voice AI that mimics human conversation in real time

Smallest.ai, founded in late 2024, raised $13 million in a Series A round led by Seligman Ventures, with parti…

TechCrunch AIJul 31, 2026

Friend AI pendant returns at $249, adds voice—double the original price

Friend, the AI pendant startup, relaunched its product with a built-in speaker that can talk to users, raising…

The Verge AIJul 30, 2026

Fish Audio raises $52M to build human-sounding AI voices

Fish Audio, an AI startup founded by former Nvidia researcher Shijia Liao, raised $52 million in seed funding…

SiliconANGLE AIJul 30, 2026

Google's Lyria 3.5 lets users edit song sections without restarting

Google released Lyria 3.5, a music generation model that produces more natural-sounding melodies, better lyric…

THE DECODERJul 29, 2026

OpenAI's GPT Transcribe improves 0.7 points but trails ElevenLabs, Google, Mistral

OpenAI released GPT Transcribe and GPT Live Transcribe, two speech recognition models via its API

THE DECODERJul 29, 2026

OpenAI launches GPT-Transcribe, speech-to-text model for audio files and live sessions

OpenAI announced GPT Transcribe, a speech-to-text model that processes completed audio files, streamed file tr…

Hacker NewsJul 29, 2026

Fish Audio raises $52M seed to build AI voice models

Palo Alto-based Fish Audio, which has over 8 million users and generates $21 million in annual recurring reven…

TechCrunch AIJul 28, 2026

Japan's Justice Ministry backs civil liability for unauthorized AI voice use

An expert committee of Japan's Justice Ministry approved a draft report Monday establishing that unauthorized…

Japan Times TechJul 28, 2026

30%+ of new podcasts are AI-generated, Listen Notes finds

Listen Notes, a podcast search database, reports that over 30% of newly created podcasts consist of AI-generat…

Hacker NewsJul 28, 2026

Deepgram integrates AWS IAM temporary delegation for SageMaker support access

Deepgram has integrated IAM temporary delegation, a new AWS IAM capability, into its support workflow for cust…

Amazon AI BlogJul 27, 2026

Ex-AWS scientist targets OpenAI, Meta with cheaper voice AI startup

Alex Smola, former distinguished scientist at Amazon, founded Boson AI to release Higgs RealTime, a speech-to-…

Fortune AIJul 27, 2026

ARIA voice-native 3D SOC platform launches under source-available license

A developer has released ARIA, a voice-controlled security operations cockpit (SOC) that runs on local hardwar…

Hacker NewsJul 26, 2026

OpenAI adds voice control to ChatGPT desktop app

OpenAI updated its ChatGPT desktop app on Thursday to support ChatGPT Voice, allowing users to speak commands…

TechCrunch AIJul 24, 2026

Claude voice mode expands to Opus, Sonnet across all platforms

Anthropic expanded Claude's voice mode to its Opus and Sonnet models (previously limited to Haiku), with the a…

THE DECODERJul 24, 2026

OpenAI brings ChatGPT Voice to desktop, hands-free task automation

OpenAI has released ChatGPT Voice on desktop and laptop computers, extending the GPT Live voice technology it…

Fortune AIJul 23, 2026

Skipper: offline AI assistant for boats, remote homes, vehicles

Skipper is an AI companion designed to work without internet, offering voice interaction, custom memory, HD vi…

Hacker NewsJul 23, 2026

Claude's voice mode expands to Opus and Sonnet models

Anthropic has made its Opus and Sonnet models available in voice mode—until now limited to the lighter Haiku m…

The Verge AIJul 23, 2026

Claude voice mode now supports three models, integrates Gmail, Slack, Notion

Anthropic updated Claude's voice mode to let users choose between Opus, Sonnet, and Haiku models, with the mod…

TechCrunch AIJul 23, 2026

Black Forest Labs releases Flux 3 with native audio video generation

German AI company Black Forest Labs released Flux 3, a multimodal foundation model that learns from images, vi…

THE DECODERJul 23, 2026

Stay ahead with AI news

Get curated AI news from 200+ sources delivered daily to your inbox. Free to use.

30,000+ monthly readers

Get Started Free

Free · takes 30 seconds · unsubscribe anytime