News & Current Events
Insights & Perspectives
AI Research
AI Products
Search
EDGENT
IQ
EDGENT
iq
About
Terms
Privacy Policy
Contact Us
EDGENT
iq
News & Current Events
Insights & Perspectives
Analytical Insights & Perspectives
Financial Sector Fortifies Against Surging AI-Powered Scams
Analytical Insights & Perspectives
Deloitte’s 2025 Outlook: Navigating Escalating AI Challenges in Human Capital
Analytical Insights & Perspectives
Salesforce Study Reveals Data Quality is Pivotal for Employee Trust in AI Adoption
Analytical Insights & Perspectives
Top Executives Sidestep Company AI Guidelines, Fueling Shadow AI Risks
Analytical Insights & Perspectives
Intel’s Evolving IP Strategy: A Calculated Shift Towards Core AI Innovation
Analytical Insights & Perspectives
Generative AI Prompts Increased Workforce Surveillance in Indian IT Sector
AI Research
AI Products
Search
EDGENT
IQ
News & Current Events
Insights & Perspectives
Analytical Insights & Perspectives
Financial Sector Fortifies Against Surging AI-Powered Scams
Analytical Insights & Perspectives
Deloitte’s 2025 Outlook: Navigating Escalating AI Challenges in Human Capital
Analytical Insights & Perspectives
Salesforce Study Reveals Data Quality is Pivotal for Employee Trust in AI Adoption
Analytical Insights & Perspectives
Top Executives Sidestep Company AI Guidelines, Fueling Shadow AI Risks
Analytical Insights & Perspectives
Intel’s Evolving IP Strategy: A Calculated Shift Towards Core AI Innovation
Analytical Insights & Perspectives
Generative AI Prompts Increased Workforce Surveillance in Indian IT Sector
AI Research
AI Products
Search
Protecting Voices from AI Cloning: E2E-VGuard’s Dual Defense Against Advanced Speech Synthesis
Step-Audio-EditX: A New Open-Source Model for Advanced Audio Editing and Text-to-Speech
ElevenLabs’ Global Ascent: A Unique Approach to AI Voice Technology
Gelina: A Unified AI Model for Synchronized Speech and Gesture Generation
ParsVoice: Unlocking High-Quality Text-to-Speech for the Persian Language
Recently Added
ControlAudio: Crafting Audio with Text, Timing, and Intelligible Speech
Read more
Advancing Speech Attribute Control with a Theoretically Grounded Autoencoder
Read more
Unlocking Interpretability: How Sparse Deepfake Detection Improves Performance and Understanding
Read more
Enhancing Text-to-Speech Accuracy with Token-Level Preference Optimization
Read more
Efficient Speech Synthesis: Introducing ECTSpeech for High-Quality One-Step Generation
Read more
TokenChain: Unlocking Efficient Speech AI with Discrete Semantic Tokens
Read more
Advancing Voice Editing and Text-to-Speech with Cross-Attentive Mamba
Read more
Automating Academic Presentations: A New Framework for Video Generation
Read more
Flamed-TTS: Advancing Zero-Shot Text-to-Speech with Efficiency and Naturalness
Read more
Fine-Grained Emotion Control in Synthetic Speech Through Feature Disentanglement
Read more
PodEval: A New Standard for Assessing AI-Generated Podcasts
Read more
Advancing Emotional Text-to-Speech with Stepwise Preference Optimization
Read more
HiStyle: Enhancing Speech Synthesis with Hierarchical Style Prediction
Read more
Emo-FiLM: Advancing Emotional Speech Synthesis with Word-Level Control
Read more
New Strategies for Enhancing Speaker Similarity in Text-to-Speech
Read more
Enhancing Speech Synthesis Stability in LLM Models Through Attention Mechanisms
Read more
Qwen3-Omni: A Unified Multimodal AI Model Excelling Across Text, Image, Audio, and Video
Read more
Automating Diverse Data Creation for Multilingual Text-to-Speech Models
Read more
Bridging the Gap: Voice Cloning and Lip-Sync for Everyday Media
Read more
LARoPE: A Smarter Way to Align Text and Speech in AI Synthesis
Read more
Unifying Speech Understanding: How FuseCodec Integrates Meaning and Context into Digital Speech
Read more
EchoX: Bridging the Semantic-Acoustic Divide in Speech-to-Speech AI
Read more
Microsoft Unveils VibeVoice: A New Era for Multi-Speaker, Long-Form AI Speech Synthesis
Read more
Tackling Evolving Audio Deepfakes with the AUDETER Dataset
Read more
New AI Method Boosts Deepfake Audio Detection Across Languages
Read more
Assessing AI Voice Interviewers for Research Data Collection
Read more
Microsoft Introduces VibeVoice: An Open-Source AI for Long-Form, Multi-Speaker Audio Generation
Read more
SwiftF0: Advancing Real-time Pitch Detection for Noisy Environments
Read more
VIBE VOICE: Advancing Long-Form, Multi-Speaker Speech Synthesis
Read more
Efficient Voice Conversion: A New Discriminator Reduces Training Time and Memory
Read more
Load more
Gen AI News and Updates
Subscribe
I have read and accepted the
Terms of Use
and
Privacy Policy
of the website and company.
- Advertisement -
What's new?
Search
Protecting Voices from AI Cloning: E2E-VGuard’s Dual Defense Against Advanced Speech Synthesis
November 11, 2025
Step-Audio-EditX: A New Open-Source Model for Advanced Audio Editing and Text-to-Speech
November 6, 2025
ElevenLabs’ Global Ascent: A Unique Approach to AI Voice Technology
November 5, 2025
Load more