Chinese InnAIO Launches AI Voice Recorder in France
InnAIO has launched two products: the TransNote AI voice recorder and the Vinabot smart photo frame....
French OVH Group to Acquire Voice AI Company Gladia
OVH Groupe has announced the initiation of exclusive negotiations to acquire Gladia, a voice-to-text...
Google Launches Gemini 3.5 Real-Time Speech Translation Model
On June 9, Google announced the launch of the Gemini 3.5 Live Translate real-time speech translation...
Microsoft Launches Windows On-Device Speech Recognition API and Aion Small Language Models
From June 2 to 3, Microsoft released updates to Windows AI APIs at Build 2026, adding an on-device s...
Google Launches Fake Call Detection for Android to Identify AI Voice Cloning Scams
Google has launched a fake call detection feature for Android to combat impersonation scams. When bo...
US-based AethexAI Raises $3 Million to Target Voice AI in Africa and the Middle East
AethexAI, a startup focused on building voice AI products for the African and Middle Eastern markets...
China's JD.com Open-Sources JoyAI-Echo Long Audio-Video Generation Framework
On June 3, JD.com launched the JoyAI-Echo long audio-video generation framework, with its code and w...
US-based Deepgram Partners with Fortanix and NVIDIA to Shift On-Premise Voice AI to Confidential Computing Deployments
Recently, US-based real-time voice AI infrastructure company Deepgram partnered with data security f...
Microsoft's MAI-Transcribe-1.5 Integrates with Foundry, 43-Language Transcription Model Completes Voice AI Workflow
On June 2, Microsoft unveiled new members of the MAI model family during Build 2026, including MAI-T...
China's Soul App Open-Sources SoulX-Transcriber, a Multi-Person Conversation Transcription Model Integrating Speaker, Timestamp, and Text Recognition
On June 3rd, the Soul App AI team (Soul AI Lab), in collaboration with the ASLP@NPU team from Northw...
China's Tencent Cloud Partners with US-Based Soniox to Promote Multilingual Voice AI
Tencent Cloud has entered into a strategic partnership with Soniox, a voice AI company headquartered...
NTT Japan Launches Multimodal Explainable AI Reasoning Framework, Visual Language Models Enter Calibration Phase for Trustworthy Output
NTT Japan recently announced the establishment of the "Rationale-Enhanced Decoding" multimodal expla...
China's Tencent Cloud Partners with US-based Soniox to Integrate Real-Time Speech Transcription into Global Communication Infrastructure
China's Tencent Cloud recently entered into a strategic partnership with Soniox, a San Francisco-bas...
China's MiniMax Launches M3 Model, Pushing AI Competition Toward Long-Task Agents with Million-Token Context
On June 1, Chinese AI company MiniMax launched its next-generation general-purpose model, MiniMax M3...
China's DeepSeek-V4-Pro API Adjusted to 1/4 of Original Pricing, Long-Term Low-Price Strategy Drives Down Large Model Inference Costs
On May 22, DeepSeek's official pricing page showed that the API price for the DeepSeek-V4-Pro model ...
Germany's DFKI and RPTU Research Team Launch Privacy Guardrail 0.2.0, Safeguarding Sensitive Data in AI Chats with Local Anonymization
The German Artificial Intelligence Research Center (DFKI) has moved the protection of sensitive data...
Anthropic Welcomes OpenAI Founding Member Andrej Karpathy, Returning to the Forefront of Large Model R&D
On May 19 local time, OpenAI founding member Andrej Karpathy updated his personal status, announcing...
xAI Launches Skills Feature Across All Platforms in the U.S., Grok Gains Permanent Cross-Conversation Memory, Evolving Towards an Automated Workspace
Elon Musk's AI company xAI announced on May 18 local time the simultaneous launch of the "Skills" fe...
US-Based Coupa Acquires UK's Rossum to Strengthen AI Document Processing, T-LLM Replaces Traditional OCR to Cover End-to-End Spend Management
Autonomous spend management platform Coupa officially announced at its Inspire 2026 conference in La...
Tencent Yuanbao Upgrades Again, Adds Smart Summary and Analysis of WeChat Chat Records
Tencent's AI assistant Yuanbao announced a feature upgrade on May 13, officially supporting smart su...
OpenAI Launches GPT-Realtime Series of Three Audio Models, Integrating GPT-5-Level Reasoning into Voice Interaction for the First Time
OpenAI has officially launched the GPT-Realtime series, comprising three real-time audio models name...
Twilio Unveils Four New Conversational Layer Capabilities at SIGNAL Conference in San Francisco, Building Persistent Memory and Cross-Channel Orchestration for Customer Interactions
On May 6, 2026, U.S. cloud communications platform Twilio officially launched four new platform capa...
Alibaba's Qwen Launches AI Voice Input on PC, Cross-Application Smart Assistant Fully Open
Alibaba's large model product "Qwen" officially launched its AI voice input feature on the PC platfo...
Baidu Library and Baidu Netdisk Jointly Release the General Agent GenFlow 4.0, With Monthly Active Users Exceeding 100 Million and Monthly Task Delivery Reaching 200 Million
On April 27, 2026, at the Baidu AI Day Open House held in Beijing, Baidu Library and Baidu Netdisk...
China's National Supercomputing Internet Launches Limited-Time Free Dialogue Service for DeepSeek-V4, with Million-Token Context Free of Charge
On April 26, China's National Supercomputing Internet platform officially launched a limited-time fr...
DeepSeek API Input Cache Price Reduced to One-Tenth of Launch Price, V4-Pro Limited-Time Offer at 0.025 CNY/Million Tokens
On April 26, DeepSeek announced an API price adjustment. The price for all API cache hits across the...
xAI Officially Launches Grok Speech-to-Text and Text-to-Speech APIs, Batch STT Processing at $0.10 per Hour
On April 17 local time, xAI announced the official launch of Speech-to-Text (STT) and Text-to-Speech...
Oracle Launches Trusted Answer Search in the US, Utilizing Vector Technology Instead of Large Language Models for Semantic Search
Oracle recently unveiled a new technology called Trusted Answer Search, which leverages vector searc...
CoreWeave Signs Multi-Year Agreement with Anthropic, Claude Model Compute Deployment to Begin This Year
On April 10, 2026, CoreWeave announced a multi-year agreement with Anthropic to provide cloud infras...
Google Meet Voice Translation Feature Officially Launches on Android and iOS Mobile Apps
On April 8 local time, Google introduced the voice translation feature to its Meet Android and iOS a...