Startups

Hive raises $85M for AI-based APIs to help moderate content, identify objects and more

Comment

Image Credits: Wenjie Dong (opens in a new window) / Getty Images

As content moderation continues to be a critical aspect of how social media platforms work — one that they may be pressured to get right, or at least do better in tackling — a startup that has built a set of data and image models to help with that, along with any other tasks that require automatically detecting objects or text, is announcing a big round of funding.

Hive, which has built a training data trove based on crowdsourced contributions from some 2 million people globally, which then powers a set of APIs that can be used to identify automatically images of objects, words and phrases — a process used not just in content moderation platforms, but also in building algorithms for autonomous systems, back-office data processing and more — has raised $85 million in funding, and the startup has confirmed that it is now valued at $2 billion.

“At the heart of what we’re doing is building AI models that can help automate work that used to be manual,” said Kevin Guo, Hive’s co-founder and CEO. “We’ve heard about RPA and other workflow automation, and that is important too but what that has also established is that there are certain things that humans should not have to do that is very structural, but those systems can’t actually address a lot of other work that is unstructured.” Hive’s models help bring structure to that other work, and Guo claims they provide “near human level accuracy.”

Glynn Capital is leading a Series D of $50 million, with General Catalyst, Tomales Bay Capital, Jericho Capital, Bain & Company and other unnamed investors participating. Hive is also confirming a Series C of $35 million led by Tomales Bay Capital in 2020 that included new strategic investments from Bain & Company and Visa. The company has now raised $121 million. 

The company has been somewhat under the radar since it was founded in 2017, in what appears to have been a pivot from founder Kevin Guo’s previous startup, a Q&A platform that was called Kiwi, which itself was a product of a project out of his time at Stanford. But since then it has quietly picked up some interesting customers, including Reddit, Yubo, Chatroulette, Omegle and Tango, along with NBCUniversal, Interpublic Group, Walmart, Visa, Anheuser-Busch InBev and more. In all it has some 100 customers and has grown more than 300% in the last year.

Hive had its start with image identification and working with companies building autonomous systems. In fact, if you talk with Guo over Zoom, chances are you’ll get a screenshot of some of that work as a background, with cars darting across Golden Gate Bridge.

These days, however, most of Hive’s activity (pardon the pun) comes around moderation, some of which includes images, but others including text and streamed audio — which is converted into text and then moderated as that would be. (The autonomous car modelling is still used as a backdrop, I believe, because it’s a little less disturbing than a content moderation image, as you can see below.)

Image Credits: Hive (opens in a new window) under a CC BY 2.0 (opens in a new window) license.

In part because it’s a very classic problem that you can imagine will be solved or helped with the use of AI, and in part because it’s such a big issue on the internet today, there are a number of other startups building platforms to help manage online abuse, including harassment and to help with content moderation.

They include the likes of Sentropy, Block Party, L1ght and Spectrum Labs, not to mention a lot of tools being built in-house by big technology companies themselves. (Instagram for example launched its latest tools to help users combat abuse in DMs just today: it built the whole thing in-house, the company told me.)

But as Kevin Guo describes it, what has set Hive apart from the crowd has been the crowd, so to speak. Over the last several years, the company has slowly been building up a trove of data by crowdsourcing feedback from some 2 million users, who get paid — either in “normal” money or Bitcoin — to go through various images and items of text in order to identify “abuse” or other things. (Bitcoin started as a fringe offering and now accounts for the majority of how contributors get paid, Guo said.)

Instagram launches tools to filter out abusive DMs based on keywords and emojis, and to block people, even on new accounts

That database in turn powers a set of APIs used by Hive’s customers to help them run their own moderation tools, or whatever workflow requires frequent and rapid identification.

Most of the language learning in the system right now is based around English and several other popular global languages such as Spanish and French. Some of the funding will be used to help expand its reach and global coverage, including into a wider set of tongues. This is also leading to a wider set of use cases for the data and technology that Hive has built.

One of these, Guo said, includes a new approach to advertising that is based around serving ads associated with something you may have just read or seen on the screen. Very GDPR-friendly because it involves absolutely no involvement of data based on you or your online browsing activities (anonymised or not), this is picking up traction with brands that initially may have come to Hive to help protect their IP or reputation management, and are now considering how they can use the tool to spread the word about themselves in more effective ways.

The possibilities for how Hive’s AI can be used in the future is part of what attracted the investment today. The focus on how it has been built in the cloud underscores that extensibility.

“Cloud computing has seen tremendous adoption in recent years, but only a small fraction of companies currently leverage cloud-based machine learning solutions,” said Charlie Friedland, principal at Glynn Capital, in a statement. “We believe cloud-hosted machine learning models will represent one of the most significant components of cloud growth in the years to come, and Hive is well-positioned as an early leader in the space.”

It’s notable to me that for now at least Hive doesn’t disclose any big technology companies among its customers. That may partly be due to NDAs, but Guo points out that their in-house activities, which include heavy doses of human involvement, have made them somewhat less willing customers up to now. That could be changing however, not just because AI tools are improving, but because of the problems that have arisen from some of the current routes, such as the run of controversial stories about social media content moderators and the traumas that they have faced.

MIT professor wants to overhaul ‘The Hype Machine’ that powers social media

In terms of future deals, those might come by way of some of Hive’s strategic backers and strategic partnerships. The company currently works with companies like Cognizant, Comscore and Bain (which is an investor), which in turn provide consulting and services to larger tech companies that have opted to outsource some of their human moderation work. Whether those human moderators shift up practices or not, chances are that tech will be playing an increasing role in the bigger process of trying to give more structure both to shaping and adhering to abuse policies.

Updated to note the two separate rounds of funding of $50 million and $35 million.

More TechCrunch

After Apple loosened its App Store guidelines to permit game emulators, the retro game emulator Delta — an app 10 years in the making — hit the top of the…

Adobe comes after indie game emulator Delta for copying its logo

Meta is once again taking on its competitors by developing a feature that borrows concepts from others — in this case, BeReal and Snapchat. The company is developing a feature…

Meta’s latest experiment borrows from BeReal’s and Snapchat’s core ideas

Welcome to Startups Weekly! We’ve been drowning in AI news this week, with Google’s I/O setting the pace. And Elon Musk rages against the machine.

Startups Weekly: It’s the dawning of the age of AI — plus,  Musk is raging against the machine

IndieBio’s Bay Area incubator is about to debut its 15th cohort of biotech startups. We took special note of a few, which were making some major, bordering on ludicrous, claims…

IndieBio’s SF incubator lineup is making some wild biotech promises

YouTube TV has announced that its multiview feature for watching four streams at once is now available on Android phones and tablets. The Android launch comes two months after YouTube…

YouTube TV’s ‘multiview’ feature is now available on Android phones and tablets

Featured Article

Two Santa Cruz students uncover security bug that could let millions do their laundry for free

CSC ServiceWorks provides laundry machines to thousands of residential homes and universities, but the company ignored requests to fix a security bug.

6 hours ago
Two Santa Cruz students uncover security bug that could let millions do their laundry for free

OpenAI’s Superalignment team, responsible for developing ways to govern and steer “superintelligent” AI systems, was promised 20% of the company’s compute resources, according to a person from that team. But…

OpenAI created a team to control ‘superintelligent’ AI — then let it wither, source says

TechCrunch Disrupt 2024 is just around the corner, and the buzz is palpable. But what if we told you there’s a chance for you to not just attend, but also…

Harness the TechCrunch Effect: Host a Side Event at Disrupt 2024

Decks are all about telling a compelling story and Goodcarbon does a good job on that front. But there’s important information missing too.

Pitch Deck Teardown: Goodcarbon’s $5.5M seed deck

Slack is making it difficult for its customers if they want the company to stop using its data for model training.

Slack under attack over sneaky AI training policy

A Texas-based company that provides health insurance and benefit plans disclosed a data breach affecting almost 2.5 million people, some of whom had their Social Security number stolen. WebTPA said…

Healthcare company WebTPA discloses breach affecting 2.5 million people

Featured Article

Microsoft dodges UK antitrust scrutiny over its Mistral AI stake

Microsoft won’t be facing antitrust scrutiny in the U.K. over its recent investment into French AI startup Mistral AI.

8 hours ago
Microsoft dodges UK antitrust scrutiny over its Mistral AI stake

Ember has partnered with HSBC in the U.K. so that the bank’s business customers can access Ember’s services from their online accounts.

Embedded finance is still trendy as accounting automation startup Ember partners with HSBC UK

Kudos uses AI to figure out consumer spending habits so it can then provide more personalized financial advice, like maximizing rewards and utilizing credit effectively.

Kudos lands $10M for an AI smart wallet that picks the best credit card for purchases

The EU’s warning comes after Microsoft failed to respond to a legally binding request for information that focused on its generative AI tools.

EU warns Microsoft it could be fined billions over missing GenAI risk info

The prospects for troubled banking-as-a-service startup Synapse have gone from bad to worse this week after a United States Trustee filed an emergency motion on Wednesday.  The trustee is asking…

A US Trustee wants troubled fintech Synapse to be liquidated via Chapter 7 bankruptcy, cites ‘gross mismanagement’

U.K.-based Seraphim Space is spinning up its 13th accelerator program, with nine participating companies working on a range of tech from propulsion to in-space manufacturing and space situational awareness. The…

Seraphim’s latest space accelerator welcomes nine companies

OpenAI has reached a deal with Reddit to use the social news site’s data for training AI models. In a blog post on OpenAI’s press relations site, the company said…

OpenAI inks deal to train AI on Reddit data

X users will now be able to discover posts from new Communities that are trending directly from an Explore tab within the section.

X pushes more users to Communities

For Mark Zuckerberg’s 40th birthday, his wife got him a photoshoot. Zuckerberg gives the camera a sly smile as he sits amid a carefully crafted re-creation of his childhood bedroom.…

Mark Zuckerberg’s makeover: Midlife crisis or carefully crafted rebrand?

Strava announced a slew of features, including AI to weed out leaderboard cheats, a new ‘family’ subscription plan, dark mode and more.

Strava taps AI to weed out leaderboard cheats, unveils ‘family’ plan, dark mode and more

We all fall down sometimes. Astronauts are no exception. You need to be in peak physical condition for space travel, but bulky space suits and lower gravity levels can be…

Astronauts fall over. Robotic limbs can help them back up.

Microsoft will launch its custom Cobalt 100 chips to customers as a public preview at its Build conference next week, TechCrunch has learned. In an analyst briefing ahead of Build,…

Microsoft’s custom Cobalt chips will come to Azure next week

What a wild week for transportation news! It was a smorgasbord of news that seemed to touch every sector and theme in transportation.

Tesla keeps cutting jobs and the feds probe Waymo

Sony Music Group has sent letters to more than 700 tech companies and music streaming services to warn them not to use its music to train AI without explicit permission.…

Sony Music warns tech companies over ‘unauthorized’ use of its content to train AI

Winston Chi, Butter’s founder and CEO, told TechCrunch that “most parties, including our investors and us, are making money” from the exit.

GrubMarket buys Butter to give its food distribution tech an AI boost

The investor lawsuit is related to Bolt securing a $30 million personal loan to Ryan Breslow, which was later defaulted on.

Bolt founder Ryan Breslow wants to settle an investor lawsuit by returning $37 million worth of shares

Meta, the parent company of Facebook, launched an enterprise version of the prominent social network in 2015. It always seemed like a stretch for a company built on a consumer…

With the end of Workplace, it’s fair to wonder if Meta was ever serious about the enterprise

X, formerly Twitter, turned TweetDeck into X Pro and pushed it behind a paywall. But there is a new column-based social media tool in town, and it’s from Instagram Threads.…

Meta Threads is testing pinned columns on the web, similar to the old TweetDeck

As part of 2024’s Accessibility Awareness Day, Google is showing off some updates to Android that should be useful to folks with mobility or vision impairments. Project Gameface allows gamers…

Google expands hands-free and eyes-free interfaces on Android