Startups

Papercup, the UK startup using AI for realistic-sounding voice translation, raises £8M funding

Comment

Image Credits: Papercup

Papercup, the U.K.-based AI startup that has developed speech technology that translates people’s voices into other languages and is already being used in the video and television industry, has raised £8 million in funding.

The round was led by LocalGlobe and Sands Capital Ventures, alongside Sky, GMG Ventures, Entrepreneur First (EF) and BDMI. Papercup says the new capital will be used to invest further into machine learning research and to expand its “human-in-the-loop” quality control functionality, which is used to improve and customise the quality of its AI-translated videos.

Meanwhile, Papercup’s existing angel investors include William Tunstall-Pedoe, the founder of Evi Technologies — the company acquired by Amazon to create Alexa — and Zoubin Ghahramani, former chief scientist and VP of AI at Uber and now part of the Google Brain leadership team.

Founded in 2017 by Jesse Shemen and Jiameng Gao while going through EF’s company builder program, Papercup is building out an AI and machine learning-based system that it says is capable of translating a person’s voice and expressiveness into other languages. Unlike a lot of text-to-speech, the startup claims the resulting voice translation is “indistinguishable” from human speech, and, perhaps uniquely, it attempts to retain the characteristics of the original speaker’s voice.

Initially, the tech is being targeted at video producers, including already being used by Sky News, Discovery and YouTube stars Yoga with Adriene, along with DIY content creators. It is pitched as a much more scalable and therefore lower-cost alternative to pure human dubbing.

“Most of the world’s video and audio content is shackled to a single language,” says Papercup co-founder and CEO Shemen. “That includes billions of hours of videos on YouTube, millions of podcast episodes, tens of thousands of classes on Skillshare and Coursera, and thousands of hours of content on Netflix. Almost every content owner is scrambling to go international, but there is yet no simple and cost-effective way to translate content beyond subtitling”.

For “deep pocketed studios,” there is of course the option to employ high-end dubbing via a professional dubbing studio and voice actors, but this is far too expensive for most content owners. And even wealthy studios are often constrained in terms of how many languages they can accommodate.

“That leaves the mid and long tail of content owners — literally 99% of all content — stranded and incapable of reaching international audiences beyond subtitling,” says Shemen, which, of course, is where Papercup comes into play. “Our aim is to generate translated voices that sound as close to the original speaker as possible”.

To do that, he says that Papercup will need to tackle four things. First up is creating “natural sounding” voices, i.e. how clear and human-like the synthetic voices sound. The second challenge is retaining emotion and pacing to reflect how the original speaker expressed themselves (think: happy, sad, angry etc.). Third is capturing the uniqueness of someone’s voice (e.g. Morgan Freeman, but in German). Lastly, the resulting translation needs the correct alignment of the audio to the video itself.

Explains Shemen: “We started off by making our voices as human-like and natural sounding as possible, where we’ve made quite a significant leap in terms of quality by honing our technology to the task, and today we have one of the best Spanish speech synthesis systems in production.

“We’re now focusing on better retainment and transfer of the original emotion and expressiveness in the original speaker across languages, and meanwhile figuring out what it is exactly that makes for quality dubbing”.

The next challenge and arguably the toughest nut to crack is “speaker adaptation,” described as capturing the uniqueness of someone’s voice. “This is the last layer of adaptation,” notes the Papercup CEO, “but it was also one of our first breakthroughs in our research. While we have models that can accomplish this, we’re focusing more of our time on emotion and expressiveness”.

That’s not to say Papercup is entirely machine-powered, even if it might be one day. The company also employs a “human-in-the-loop” process to make corrections and adjustments to the translated audio track. This includes correcting for any speech recognition or machine translation errors that come up, making adjustments to the timings of the audio, as well as enforcing emotions (e.g. happy, sad) and changing the speed of the generated voice.

How much human-in-the-loop is required depends on the type of content and priorities of the content owners, i.e. how realistic or perfect they need the resulting video to be. In other words, it isn’t a zero-sum game, as good enough will be more than enough for a swathe of content owners at scale.

Meet the startups that pitched at EF’s 9th Demo Day in London

Asked about the technology’s beginnings, Shemen says Papercup started with research conducted by co-founder and CTO Jiameng Gao “who is incredibly smart and oddly obsessed with speech processing”. Gao completed two Masters at University of Cambridge (in machine learning and speech language technology) and wrote a thesis on speaker adaptive speech processing. It was at Cambridge that he realised that something like Papercup was possible.

“When we started working together at Entrepreneur First at the end of 2017, we built our initial prototype systems that showed that this technology was even possible despite there being no precedent for it,” says Shemen. “Based on early conversations, the demand was clearly overwhelming for what we were building — it was just a function of actually building something that could be used in a production environment”.

More TechCrunch

The change would see Instagram becoming more like the free version of YouTube, which requires users to view ads before and in the middle of watching videos.

Instagram confirms test of ‘unskippable’ ads

Commerce platform Shopify has acquired Checkout Blocks, allowing Shopify Plus merchants to make no-code customizations in their checkout to enhance customer experience and potentially boost sales.  Checkout Blocks, which debuted…

Shopify acquires Checkout Blocks, a checkout customization app

After the Digital Markets Act (DMA) forced Apple to allow third-party app stores for iOS in Europe, several developers have launched alternative stores, like the AltStore and MacPaw’s Setapp (currently…

Aptoide launches its alternative iOS game store in the EU

Time is relentless and, right now, it’s no friend to procrastination-prone early-stage startup founders. The application window for Startup Battlefield 200 (SB 200) at TechCrunch Disrupt 2024 slams shut in…

One week left: Apply to TC Disrupt Startup Battlefield 200

Cloudera, the once high flying Hadoop startup, raised $1 billion and went public in 2018 before being acquired by private equity for $5.3 billion 2021. Today, the company announced that…

Cloudera acquires Verta to bring some AI chops to its data platform

The global spend management sector is experiencing a tailwind of sorts. North America is arguably the biggest market in this space, but spend management companies have seen demand rise across…

Spend management startup SiFi raises $10M to grow further in Saudi Arabia

Neural Concept lets designers model how components will perform before they can be manufactured.

Swiss startup Neural Concept raises $27M to cut EV design time to 18 months

The StrictlyVC roadtrip continues! Coming off of sold-out events in London, Los Angeles, and San Francisco, we’re heading to Washington, D.C. for a cozy-vc-packed, evening at the Woolly Mammoth Theatre…

Don’t miss StrictlyVC in DC next week

X will now allow users to post consensually produced NSFW content as long as it is prominently labeled as such.

X tweaks rules to formally allow adult content

Ashby consolidates existing talent acquisition tools and leans heavily on AI to automate the more repetitive steps in the recruitment pipeline.

Ashby injects recruiting with a dose of AI

Spotify has announced it’s hiking subscriptions for customers in the U.S., the second such price increase in the space of a year. The music-streaming giant reports that premium pricing will…

Spotify to increase premium pricing in the US to $11.99 per month

Monzo has announced its 2024 financial results, revealing its first full-year pre-tax profit. The company also confirmed that it’s in the early stages of expanding into the broader European market…

UK neobank Monzo reports first full (pre-tax) profit, prepares for EU expansion with Dublin hub

Featured Article

Inside Apple’s efforts to build a better recycling robot

Last week, TechCrunch paid a visit to Apple’s Austin, Texas manufacturing facilities. Since 2013, the company has built its Mac Pro desktop about 20 minutes north of downtown. The 400,000-square-foot facility sits in a maze of industry parks, a quick trip south from the company’s in-progress corporate campus. In recent years, the capital city has…

8 hours ago
Inside Apple’s efforts to build a better recycling robot

Early attempts at making dedicated hardware to house artificial intelligence smarts have been criticized as, well, a bit rubbish. But here’s an AI gadget-in-the-making that’s all about rubbish, literally: Finnish…

Binit is bringing AI to trash

Temasek has previously invested in Lenskart, and this new funding follows a $500 million investment by the Abu Dhabi Investment Authority last year.

Temasek, Fidelity buy $200M stake in Lenskart at $5B valuation

Less than one year after its iOS launch, French startup ten ten has gone viral with a walkie talkie app that allows teens to send voice messages to their close…

French startup ten ten reinvents the walkie-talkie

Featured Article

Unicorn-rich VC Wesley Chan owes his success to a Craigslist job washing lab beakers

While all of Wesley Chan’s success has been well-documented over the years, his personal journey…not so much. Chan spoke to TechCrunch about the ways his life impacts how he invests in startups.

1 day ago
Unicorn-rich VC Wesley Chan owes his success to a Craigslist job washing lab beakers

Presumptive Republican presidential nominee Donald Trump now has an account on the short-form video app that he once tried to ban. Trump’s TikTok account, which launched on Saturday night, features…

Trump takes off on TikTok

With fewer than 400,000 inhabitants, Iceland receives more than its fair share of tourists — and of venture capital.

Iceland’s startup scene is all about making the most of the country’s resources

Kobo put out a handful of new e-readers a few weeks back: color versions of the excellent Libra 2 and Clara, as well as an updated monochrome version of the…

Kobo’s new e-readers are a sidegrade most can skip (with one exception)

In an interview at his home near Reykjavík, the entrepreneur-turned-VC shared thoughts on his ventures and the journey that led him from Unity to climate tech, a homecoming of sorts.

Unity co-founder David Helgason’s next act: Gaming the climate crisis

Welcome back to TechCrunch’s Week in Review — TechCrunch’s newsletter recapping the week’s biggest news. Want it in your inbox every Saturday? Sign up here. Over the past eight years,…

Fisker collapsed under the weight of its founder’s promises

What is AI? We’ve put together this non-technical guide to give anyone a fighting chance to understand how and why today’s AI works.

WTF is AI?

President Joe Biden has vetoed H.J.Res. 109, a congressional resolution that would have overturned the Securities and Exchange Commission’s current approach to banks and crypto. Specifically, the resolution targeted the…

President Biden vetoes crypto custody bill

Featured Article

Industries may be ready for humanoid robots, but are the robots ready for them?

How large a role humanoids will play in that ecosystem is, perhaps, the biggest question on everyone’s mind at the moment.

2 days ago
Industries may be ready for humanoid robots, but are the robots ready for them?

VCs are clamoring to invest in hot AI companies, and willing to pay exorbitant share prices for coveted spots on their cap tables. Even so, most aren’t able to get…

VCs are selling shares of hot AI companies like Anthropic and xAI to small investors in a wild SPV market

The fashion industry has a huge problem: Despite many returned items being unworn or undamaged, a lot, if not the majority, end up in the trash. An estimated 9.5 billion…

Deal Dive: How (Re)vive grew 10x last year by helping retailers recycle and sell returned items

Tumblr officially shut down “Tips,” an opt-in feature where creators could receive one-time payments from their followers.  As of today, the tipping icon has automatically disappeared from all posts and…

You can no longer use Tumblr’s tipping feature 

Generative AI improvements are increasingly being made through data curation and collection — not architectural — improvements. Big Tech has an advantage.

AI training data has a price tag that only Big Tech can afford

Keeping up with an industry as fast-moving as AI is a tall order. So until an AI can do it for you, here’s a handy roundup of recent stories in the world…

This Week in AI: Can we (and could we ever) trust OpenAI?