AIToday

Catch up on AI in one minute, every morning.

Start free30,000+ monthly readers

Latest AI News

Earlier

Komatsu tests voice AI for equipment maintenance

SORABITO and Eleven Labs have built an interactive voice AI system that diagnoses equipment faults from verbal…

Top Companies AIJul 22, 2026

Deezer: AI tracks now over 50% of daily uploads

Music streaming platform Deezer announced that AI-generated music now represents more than 50% of daily upload…

TechCrunch AIJul 21, 2026

Sony sues Udio over 30,000 copyrighted songs

Sony Music Entertainment filed a new lawsuit in New York against AI music generator Udio on Monday, claiming i…

The Verge AIJul 21, 2026

Suno AI music finally doesn't sound terrible—here's why

Artist 1010Benja released a song called "Semiramis' Dream" on his latest EP using Suno generative AI, but in a…

The Verge AIJul 19, 2026

Suno trained AI on millions of songs scraped from YouTube, Genius, Deezer

A hacking incident exposed that Suno, an AI music generator, scraped millions of songs and lyrics from online…

The Verge AIJul 15, 2026

Suno AI music tool hacked; source code reveals YouTube scraping

The AI music generator Suno was hacked via a supply chain attack that exposed source code showing the platform…

TechCrunch AIJul 15, 2026

Spotify adds chat feature for Premium subscribers to control music with voice and text

Spotify is rolling out a conversational AI interface for Premium subscribers, allowing them to control playbac…

THE DECODERJul 15, 2026

Hugging Face launches voice AI benchmark measuring human perception, not just speed

Hugging Face introduced Real World VoiceEQ, a benchmark that evaluates more than 40 voice models across 15+ di…

Hugging Face BlogJul 15, 2026

Telnyx launches AI phone agent for real-time price quotes

Telnyx released an open-source Python application (102 lines) that uses Telnyx Call Control and Llama 3.3 70B…

Hacker NewsJul 15, 2026

ScienceSoft builds HIPAA-compliant AI voice scheduler on AWS

AWS Partner ScienceSoft has built a HIPAA-compliant AI voice scheduling assistant using Amazon Nova Sonic and…

Amazon AI BlogJul 14, 2026

Apple Expands Siri Into All-Purpose iPhone Assistant

Apple released iOS 27 as a public beta, making the revamped "Siri AI" available to iPhone users for the first…

WIRED AIJul 14, 2026

OpenAI launches GPT-Live voice models, plans GPT-5.6 broad release

OpenAI introduced GPT-Live, a family of voice-optimized AI models powering ChatGPT's voice mode

SiliconANGLE AIJul 10, 2026

OpenAI upgrades voice mode; Grok 4.5 launches

OpenAI released an upgraded voice mode described as a step change, and Grok 4.5 became available

LessWrong AIJul 10, 2026

Paris AI voice startup Gradium raises $100M seed, expands to Bay Area

Gradium, a Paris-based startup building voice AI models, closed its seed round at $100 million(約160億円) total…

TechCrunch AIJul 9, 2026

OpenAI upgrades ChatGPT voice with full-duplex model that listens and speaks simultaneously

OpenAI is rolling out GPT-Live-1, a new voice model for ChatGPT that can listen and speak at the same time, re…

The Verge AIJul 8, 2026

OpenAI releases full-duplex voice models for longer, more natural ChatGPT conversations

OpenAI released two new conversational voice models—GPT-Live-1 and GPT-Live-1 mini—that can speak and listen s…

TechCrunch AIJul 8, 2026

Cohere releases open-source Arabic speech-to-text model

Cohere released Cohere Transcribe Arabic, a 2-billion-parameter open-source model for Arabic speech recognitio…

THE DECODERJul 7, 2026

Speech Recognition Gets 2.8x Faster With Simpler Decoder Change

Researchers found that speech recognition systems waste compute processing silence instead of speech

Daily Dose of Data ScienceJul 2, 2026

Revin AI agents now live across Vertex roofing platform

Revin, an AI voice and SMS platform, launched its agents across Vertex Service Partners' portfolio of 27 regio…

Top Companies AIJul 1, 2026

Hugging Face and Cerebras demo real-time voice AI with low latency

Hugging Face and Cerebras demonstrated a speech-to-speech AI pipeline that combines open-source models—Nvidia'…

Hugging Face BlogJul 1, 2026

Netflix uses AI-generated Gene Wilder voice for Wonka reality show

Netflix is premiering Wonka's The Golden Ticket on September 23rd, a reality competition based on the fictiona…

The Verge AIJun 30, 2026

Jamendo sues Suno over alleged unauthorized use of music in AI training

Jamendo SA, a subsidiary of Winamp Group, filed a federal lawsuit in Massachusetts against Suno, Inc., allegin…

Yahoo Finance AIJun 30, 2026

AI Voice Studio launches a free-to-start tool that lets creators generate natural voiceovers in 70+ languages without recording equipment or waiting for voice talent.

AI Voice Studio is now available, allowing creators to convert scripts into studio-quality voiceovers across 7…

Hacker NewsJun 20, 2026

Whissle Gateway lets businesses run voice AI—including speech recognition, speaker identification, and sales coaching analysis—entirely on their own servers in a 500MB Docker container.

Whissle released a containerized voice AI system that performs automatic speech recognition (ASR), text-to-spe…

Hacker NewsJun 13, 2026

Google releases Gemini 3.5 Live Translate, a real-time audio translation model supporting over 70 languages

Gemini 3.5 Live Translate is available now for developers through the Gemini Live API and Google AI Studio, as…

THE DECODERJun 9, 2026

ElevenLabs signs Memorandum of Understanding with UK's Department for Science, Innovation and Technology to deploy voice AI in public services; company doubles UK headcount to 200 and expands London headquarters.

ElevenLabs and the UK's Department for Science, Innovation and Technology (DSIT) agreed to collaborate on thre…

ElevenLabs BlogJun 8, 2026

Apple introduces Siri AI with multi-step conversational abilities and Google-powered model update, rolling out this fall

At its Worldwide Developers Conference, Apple announced 'Siri AI'—an updated voice assistant coming in OS upda…

Ars Technica AIJun 8, 2026

Researchers develop three-billion-parameter model that listens continuously to audio and decides every 0.4 seconds whether to speak, handling translation, transcription, and sound recognition simultaneously

Audio-Interaction, created by researchers from China, Hong Kong, and Singapore, processes continuous audio str…

THE DECODERJun 6, 2026

AI voice customer service struggles with recognition and user trust despite technological progress

ElevenLabs demonstrated a voice-powered robot at a New York pop-up that took coffee orders and prepared drinks…

Semafor TechJun 5, 2026

AethexAI, founded by ex-Goldman and ex-Meta executives, raises $3 million to build voice AI for Africa and the Middle East

AethexAI raised $3 million in pre-seed funding led by 4DX Ventures, with participation from Enza Capital, Dorm…

TechCrunch AIJun 3, 2026

AI music startup Suno raises $400 million at $5.4 billion valuation, doubling its worth in seven months

Suno raised $400 million at a $5.4 billion valuation, double its valuation from seven months ago

THE DECODERJun 3, 2026

SoundHound AI stock rises 5.7% after Snowflake earnings show AI driving more platform consumption, not replacing software

Snowflake reported that AI accounts on its platform jumped from 9,100 to 13,600 in a single quarter, product r…

Yahoo Finance AIMay 29, 2026

AI voice generator market projected to reach USD 20.71 billion by 2031, growing at 30.7% CAGR from USD 4.16 billion in 2025

The AI voice generator market is set to achieve a compound annual growth rate (CAGR) of 30.7% over the forecas…

Yahoo Finance AIMay 27, 2026

Suno users report listening almost exclusively to their own AI-generated music instead of songs by professional artists

Users in the Suno subreddit describe a pattern of consuming primarily their own AI-generated music over tradit…

The Verge AIMay 26, 2026

Stripe developer relations leaders James Beswick and Peter Epstein are authors of an article titled 'You can't whisper at an AI agent'

James Beswick leads the Stripe Developer Relations team and was previously a Developer Advocacy leader at AWS

Hacker NewsMay 24, 2026

Amazon SageMaker AI and vLLM enable real-time speech-to-text via bidirectional streaming, starting November 2025

Starting November 2025, Amazon SageMaker AI supports bidirectional streaming for real-time inference, allowing…

Amazon AI BlogMay 20, 2026

Five minutes each morning. That's the whole AI news habit.

The day's essentials from 200+ sources, delivered to Email, LINE, or Slack. Free, always.

30,000+ monthly readers
Get Started Free

Stability AI launches Stable Audio 3.0 with models generating up to six-minute tracks, three variants released as open weights

Stability AI unveiled Stable Audio 3.0, a family of four audio models

THE DECODERMay 20, 2026

Stability AI releases Stability Audio 3.0 with models capable of generating six-minute songs

Stability AI released four new audio models under the Stability Audio 3.0 name: small SFX (459M parameters), s…

TechCrunch AIMay 20, 2026

ElevenLabs launches professor access program and recreates Albert Einstein's voice for interactive learning

ElevenLabs is offering professors free access to its Pro tier and the ability to provide time-bound access to…

ElevenLabs BlogMay 19, 2026

Thinking Machines Lab releases its first AI model with native interactivity, claiming it outperforms OpenAI's GPT-Realtime-2 and Google's Gemini Live on interaction quality

Thinking Machines Lab, founded in February 2025 by Mira Murati and other former OpenAI researchers, published…

THE DECODERMay 12, 2026

Rivian's AI-powered voice assistant rolls out to vehicle owners via software update

Rivian is releasing its Rivian Assistant to all compatible Gen 1 and Gen 2 vehicle owners who subscribe to Con…

The Verge AIMay 12, 2026

Lightweight voice gender classifier for European languages released: 0.64 MB ONNX model runs in ~4 ms on CPU

A Bi-LSTM gender classifier model (166K parameters, 0.64 MB) designed for real-time voice AI pipelines support…

Hacker NewsMay 12, 2026

Vapi raises $50 million Series B at $500 million valuation after Amazon Ring chose it over 40 rivals to handle all inbound customer calls

Amazon Ring evaluated more than 40 AI voice vendors before selecting Vapi to route 100% of its inbound calls t…

TechCrunch AIMay 12, 2026

Wispr Flow expands India operations with Hinglish voice model and aggressive hiring as the country becomes its second-largest market

Wispr Flow, a Bay Area-headquartered startup building AI-powered voice input software, began beta testing a Hi…

TechCrunch AIMay 10, 2026

TypeWhisper 1.4 release candidate adds unified workflows, system-wide dictation, and support for ten transcription engines on macOS

TypeWhisper 1.4 is the current release-candidate line for macOS, featuring system-wide dictation, file transcr…

Hacker NewsMay 9, 2026

Dikaletus: open-source meeting agent script automates audio recording, transcription, and note generation using Mistral AI

The Meeting Agent is a script that records audio from microphone and speaker outputs using PulseAudio and FFmp…

Hacker NewsMay 9, 2026

OpenAI launches GPT-Realtime-2, GPT-Realtime-Translate, and GPT-Realtime-Whisper voice models in Realtime API with expanded context and reasoning controls

OpenAI released three streaming audio models: GPT-Realtime-2 (a native speech-to-speech model for voice agents…

Latent SpaceMay 8, 2026

OpenAI releases GPT-Realtime-2 voice model with reasoning capabilities, plus live translation and transcription models

OpenAI shipped three new voice models: GPT-Realtime-2 (for reasoning and real-time conversation), GPT-Realtime…

THE DECODERMay 7, 2026

Google Home gets Gemini 3.1 voice assistant with improved reasoning for multi-step commands and expands Ask Home chatbot to web interface

Google rolled out Gemini 3.1 to Google Home smart speakers, initially available to early access users

Ars Technica AIMay 5, 2026

ElevenLabs reveals new investors in $500 million Series D, including BlackRock, NVIDIA, and celebrity backers; company surpassed $500 million in ARR and reached $11 billion valuation

ElevenLabs announced additional investors in its Series D fundraise first announced in February, including ins…

TechCrunch AIMay 5, 2026

InterviewDen launches free AI voice mock interviews for software engineering, consulting, banking, and quantitative finance roles.

InterviewDen offers live voice and text mock interviews that listen, speak, and grade responses

Hacker NewsApr 29, 2026

AWS publishes guide for migrating text agents to voice assistants using Amazon Nova 2 Sonic

AWS published a blog post exploring how to convert a traditional text agent into a conversational voice assist…

Amazon AI BlogApr 28, 2026

Arietta Voice: open-source local-first framework for building wake-word-driven voice assistants on Apple Silicon Macs

Arietta Voice combines local speech-to-text (Moonshine), text-to-speech (Kokoro), turn detection (Silero VAD a…

Hacker NewsApr 28, 2026

PAVO: an 85,041-parameter router for voice pipelines cuts P95 latency 10.3% and energy 71% versus fixed-cloud, trained via PPO in 106 seconds on a 50,000-turn benchmark.

Researchers at University of Pennsylvania and Google released PAVO-Bench, a 50,000-turn voice interaction data…

Hacker NewsApr 28, 2026

VoiceGoat: Open-source vulnerable voice agent platform for security training released on GitHub

VoiceGoat is a modular platform designed for security practitioners to practice exploiting voice-based AI syst…

Hacker NewsApr 28, 2026

Microsoft VibeVoice ASR integrated into Hugging Face Transformers; voice AI framework now supports 60-minute speech recognition in single pass

VibeVoice-ASR, a speech-to-text model, is now part of a Transformers release and available directly through th…

Hacker NewsApr 28, 2026

Canonical plans to add AI features to Ubuntu Linux throughout 2026, prioritizing model transparency and local inference

Jon Seager, VP of engineering at Canonical, shared a blog post on Monday detailing plans to add AI features to…

The Verge AIApr 27, 2026

ComfyUI raises $30M at $500M valuation, giving creators manual control over AI-generated images, videos, and audio

ComfyUI, a startup building AI creation tools, closed a $30 million funding round that values the company at $…

TechCrunch AIApr 24, 2026

Over 50% of American workers now use AI daily — but most are experimenting quietly, not following influencer tutorials

A LessWrong author surveyed their own AI usage patterns and found themselves using AI assistance for hours eve…

LessWrong AIApr 21, 2026

New text-to-speech platform TTS.ai launches with minimal early traction on Hacker News

TTS.ai is a new text-to-speech service that was shared on Hacker News

Hacker NewsApr 18, 2026

Developer launches PrivaKit, a privacy-first AI workspace that runs entirely in the browser using WebGPU to eliminate cloud API risks for sensitive document processing.

PrivaKit uses transformers.js, Whisper, and WebGPU to perform AI tasks like transcription, OCR, and image proc…

Hacker NewsApr 18, 2026

Sony Music sues Udio for allegedly scraping YouTube music to train its AI music generation model without licensing.

Sony Music filed a lawsuit against Udio, claiming the AI music startup used stream ripping to extract audio fr…

Hacker NewsApr 15, 2026

Google's Gemini 3.1 Flash now includes native text-to-speech capabilities for faster audio generation

Gemini 3.1 Flash model adds integrated text-to-speech (TTS) functionality

Simon Willison's WeblogApr 15, 2026

Google rolls out Gemini 3.1 Flash TTS across its product suite, bringing more expressive and natural AI-generated speech capabilities to users.

Gemini 3.1 Flash TTS is now available across multiple Google products and services

Google AI BlogApr 15, 2026

Developer creates a 24/7 AI-powered YouTube livestream that generates unique songs every few minutes with lyrics describing the current time.

Fully automated system using Suno's API generates new AI songs in different genres every few minutes, with lyr…

Hacker NewsApr 14, 2026

Video essay explores concerns about AI music generation tools like Suno and their potential negative impacts on the music industry and creators

Suno is an AI music generation platform that raises questions about the future of music creation and industry…

Hacker NewsApr 12, 2026

OpenAI's ChatGPT voice mode reportedly uses a less capable underlying model than the text-based version

ChatGPT's voice interface appears to be powered by a weaker AI model compared to the standard text-based ChatG…

Simon Willison's WeblogApr 10, 2026

ByteDance introduces Seeduplex, a full-duplex voice AI system enabling natural two-way conversations without turn-taking delays.

Seeduplex represents ByteDance's advancement in conversational AI, allowing simultaneous speaking and listenin…

Hacker NewsApr 10, 2026

Discrete speech units struggle to preserve lexical tone information in Mandarin and Yorùbá despite SSL models encoding it naturally

Self-supervised learning (SSL) models successfully capture tone information in their latent representations, b…

arXiv cs.CLApr 10, 2026

New speech recognition benchmark reveals academic tests miss real-world challenges like custom vocabulary that matters most to users

Contextual Earnings-22 dataset created to address gap between academic benchmarks and actual industrial speech…

arXiv cs.CLApr 10, 2026

ElevenLabs expands its voice AI capabilities with local deployment options for enterprise customers

ElevenLabs now offers on-premise deployment, allowing enterprises to run voice AI solutions within their own i…

ElevenLabs BlogApr 9, 2026

Microsoft executives acknowledge voice AI technology requires significant advancement to support the company's enterprise artificial intelligence strategy.

Microsoft identifies improved voice understanding as a critical gap in its AI development roadmap

Yahoo Finance AIApr 9, 2026

Stay ahead with AI news

Get curated AI news from 200+ sources delivered daily to your inbox. Free to use.

30,000+ monthly readers

Get Started Free

Free · takes 30 seconds · unsubscribe anytime