Close Menu
World Forbes – Business, Tech, AI & Global Insights
  • Home
  • AI
  • Billionaires
  • Business
  • Cybersecurity
  • Education
    • Innovation
  • Money
  • Small Business
  • Sports
  • Trump
What's Hot

How and when to clean your reusable water bottle

November 8, 2025

Struggling families need help feeding pets as SNAP payments in doubt

November 8, 2025

Women in Mexico find safety in a feminist rideshare network

November 8, 2025
Facebook X (Twitter) Instagram
Trending
  • How and when to clean your reusable water bottle
  • Struggling families need help feeding pets as SNAP payments in doubt
  • Women in Mexico find safety in a feminist rideshare network
  • More Pakistani women are joining the country’s firefighters
  • Musk’s Net Worth Drops $10 Billion—And Tesla Shares Fall—Here’s Why
  • Here’s what to know about a study that raises questions about melatonin use and heart health
  • Trump’s Bungled Bet On Bitcoin Is Costing Him Bigtime
  • A Startup Was Their First-Ever Job—Now They’re The World’s Youngest Self Made Billionaires
World Forbes – Business, Tech, AI & Global InsightsWorld Forbes – Business, Tech, AI & Global Insights
Saturday, November 8
  • Home
  • AI
  • Billionaires
  • Business
  • Cybersecurity
  • Education
    • Innovation
  • Money
  • Small Business
  • Sports
  • Trump
World Forbes – Business, Tech, AI & Global Insights
Home » Two undergrads built an AI speech model to rival NotebookLM
AI

Two undergrads built an AI speech model to rival NotebookLM

By adminApril 22, 2025No Comments3 Mins Read
Facebook Twitter Pinterest LinkedIn Tumblr WhatsApp Telegram Email
Share
Facebook Twitter LinkedIn Pinterest Email
Post Views: 98


A pair of undergrads, neither with extensive AI expertise, say that they’ve created an openly available AI model that can generate podcast-style clips similar to Google’s NotebookLM.

The market for synthetic speech tools is vast and growing. ElevenLabs is one of the largest players, but there’s no shortage of challengers (see PlayAI, Sesame, and so on). Investors believe that these tools have immense potential. According to PitchBook, startups developing voice AI tech raised over $398 million in VC funding last year.

Toby Kim, one of the Korea-based co-founders of Nari Labs, the group behind the newly released model, said that he and his fellow co-founder started learning about speech AI three months ago. Inspired by NotebookLM, they wanted to create a model that offered more control over generated voices and “freedom in the script.”

Kim says they used Google’s TPU Research Cloud program, which provides researchers with free access to the company’s TPU AI chips, to train Nari’s model, Dia. Weighing in at 1.6 billion parameters, Dia can generate dialogue from a script, letting users customize speakers’ tones and insert disfluencies, coughs, laughs, and other nonverbal cues.

Parameters are the internal variables models use to make predictions. Generally, models with more parameters perform better.

Available from the AI dev platform Hugging Face and GitHub, Dia can run on most modern PCs with at least 10GB of VRAM. It generates a random voice unless prompted with a description of an intended style, but it can also clone a person’s voice.

In TechCrunch’s brief testing of Dia through Nari’s web demo, Dia worked quite well, uncomplainingly generating two-way chats about any subject. The quality of the voices seems competitive with other tools out there, and the voice cloning function is among the easiest this reporter has tried.

Here’s a sample:

Like many voice generators, Dia offers little in the way of safeguards, however. It’d be trivially easy to craft disinformation or a scammy recording. On Dia’s project pages, Nari discourages abuse of the model to impersonate, deceive, or otherwise engage in illicit campaigns, but the group says it “isn’t responsible” for misuse.

Nari also hasn’t disclosed which data it scraped to train Dia. It’s possible Dia was developed using copyrighted content — a commenter on Hacker News notes that one sample sounds like the hosts of NPR’s “Planet Money” podcast. Training models on copyrighted content is a widespread but legally dubious practice. Some AI companies claim that fair use shields them from liability, while rights holders assert that fair use doesn’t apply to training.

In any event, Kim says Nari’s plan is to create a synthetic voice platform with a “social aspect” on top of Dia and larger, future models. Nari also intends to release a technical report for Dia, and to expand the model’s support to languages beyond English.



Source link

Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
admin
  • Website

Related Posts

After Klarna, Zoom’s CEO also uses an AI avatar on quarterly call

May 23, 2025

Anthropic CEO claims AI models hallucinate less than humans

May 22, 2025

Anthropic’s latest flagship AI sure seems to love using the ‘cyclone’ emoji

May 22, 2025

A safety institute advised against releasing an early version of Anthropic’s Claude Opus 4 AI model

May 22, 2025

Anthropic’s new AI model turns to blackmail when engineers try to take it offline

May 22, 2025

Meta adds another 650 MW of solar power to its AI push

May 22, 2025
Add A Comment
Leave A Reply

Don't Miss
Billionaires

Musk’s Net Worth Drops $10 Billion—And Tesla Shares Fall—Here’s Why

November 7, 2025

ToplineTesla shares declined more than 3% on Friday, cutting CEO Elon Musk’s fortune by $10…

Trump’s Bungled Bet On Bitcoin Is Costing Him Bigtime

November 7, 2025

A Startup Was Their First-Ever Job—Now They’re The World’s Youngest Self Made Billionaires

November 7, 2025

Meet The Former Journalist Giving Away Billions

November 7, 2025
Our Picks

How and when to clean your reusable water bottle

November 8, 2025

Struggling families need help feeding pets as SNAP payments in doubt

November 8, 2025

Women in Mexico find safety in a feminist rideshare network

November 8, 2025

More Pakistani women are joining the country’s firefighters

November 7, 2025

Subscribe to Updates

Subscribe to our newsletter and never miss our latest news

Subscribe my Newsletter for New Posts & tips Let's stay updated!

About Us
About Us

Welcome to World-Forbes.com
At World-Forbes.com, we bring you the latest insights, trends, and analysis across various industries, empowering our readers with valuable knowledge. Our platform is dedicated to covering a wide range of topics, including sports, small business, business, technology, AI, cybersecurity, and lifestyle.

Our Picks

After Klarna, Zoom’s CEO also uses an AI avatar on quarterly call

May 23, 2025

Anthropic CEO claims AI models hallucinate less than humans

May 22, 2025

Anthropic’s latest flagship AI sure seems to love using the ‘cyclone’ emoji

May 22, 2025

Subscribe to Updates

Subscribe to our newsletter and never miss our latest news

Subscribe my Newsletter for New Posts & tips Let's stay updated!

Facebook X (Twitter) Instagram Pinterest
  • Home
  • About Us
  • Advertise With Us
  • Contact Us
  • DMCA Policy
  • Privacy Policy
  • Terms & Conditions
© 2025 world-forbes. Designed by world-forbes.

Type above and press Enter to search. Press Esc to cancel.