Startups

Snorkel AI scores $35M Series B to automate data labeling in machine learning

Comment

Anna Bassi / EyeEm
Image Credits: Anna Bassi / EyeEm / Getty Images

One of the more tedious aspects of machine learning is providing a set of labels to teach the machine learning model what it needs to know. Snorkel AI wants to make it easier for subject matter experts to apply those labels programmatically, and today the startup announced a $35 million Series B.

It also announced a new tool called Application Studio that provides a way to build common machine learning applications using templates and predefined components.

Lightspeed Venture Partners led the round with participation from previous investors Greylock, GV, In-Q-Tel and Nepenthe Capital. New investors Walden and BlackRock also joined in. The startup reports that it has now raised $50 million.

Company co-founder and CEO Alex Ratner says that data labeling remains a huge challenge and roadblock to moving machine learning and artificial intelligence forward inside a lot of industries because it is costly, labor-intensive and hard for the subject experts to carve out the time to do it.

“The not so hidden secret about AI today is that in spite of all the technological and tooling advancements, roughly 80 to 90% of the cost and time for an average AI project goes into just manually labeling and collecting and relabeling this training data,” he said.

He says that his company has developed a solution to simplify this process to make it easier for subject experts to programmatically add the labels, a process he says decreases the time and effort required to apply labels in a pretty dramatic way from months to hours or days, depending on the complexity of the data.

As the company has developed this methodology, customers have been asking for help in the next step of the machine learning process, which is taking that training data and the model and building an application. That’s where the Application Studio comes in. It could be a contract classifier at a bank or a network anomaly detector at a telco and it helps companies take that next step after data labeling.

“It’s not just about how you programmatically label the data, it’s also about the models, the preprocessors, the post processors, and so we’ve made this now accessible in a kind of templated and visual no-code interface,” he said.

DataRobot is acquiring Paxata to add data prep to machine learning platform

The company’s products are based on research that began at the Stanford AI Lab in 2015. The founders spent four years in the research phase before launching Snorkel in 2019. Today, the startup has 40 employees. Ratner recognizes the issues that the technology industry has had from a diversity perspective and says he has made a conscious effort to build a diverse and inclusive company.

“What I can say is that we tried to prioritize it at a company level, the full team level and at a board level from day one, and to also put action behind that. So we’ve been working with external firms for internal training and audits and strategy around DEI, and we’ve made pipeline diversity a non-negotiable requirement of any of our contracts with recruiting firms,” he said.

Ratner also recognizes that automation can hard code bias into machine learning models, and he’s hopeful that by simplifying the labeling process, it can make it much easier to detect bias when it happens.

“If you start with a dozen or two dozen of what we call labeling functions in Snorkel, you still need to be vigilant and proactive about trying to detect bias, but it’s easier to audit what taught your model to change it by just going back and looking at a couple of hundred lines of code.”

How artificial intelligence will be used in 2021

More TechCrunch

Long-time Android Engineering VP Dave Burke said today that he is stepping down from the role. Burke, who spent 14 years building Android, is not leaving Alphabet and is exploring…

Android Engineering VP Dave Burke steps down, as he explores “AI/bio” roles within the company

When Jordan Nathan launched his DTC nontoxic cookware company, Caraway, in 2019, he knew he was not the only founder trying to sell a new brand of pots and pans…

Why being the last company to launch in a category can pay off

Out of an abundance of caution, the car took two minutes to turn a corner.

This humanoid robot can drive cars — sort of

There has been a silly amount of drama in the run-up to Tesla‘s annual shareholder meeting on Thursday. The company is set to hold a vote on “re-ratifying” the $56…

Ahead of Tesla’s big shareholder vote, let’s re-read the judge’s opinion that got us here

To give users more control over the contacts an app can and cannot access, the permissions screen has two stages.

iOS 18 cracks down on apps asking for full address book access

The push to produce a robotic intelligence that can fully leverage the wide breadth of movements opened up by bipedal humanoid design has been a key topic for researchers.

Generative AI takes robots a step closer to general purpose

A TechCrunch review of LinkedIn data found that Ford has built this team up to around 300 employees over the last year.

Ford’s secretive, low-cost EV team is growing with talent from Rivian, Tesla and Apple

The most critical systems of our modern world rely on GPS, from aviation and road networks to emergency and disaster response, from precision farming and power grids to weather forecasting…

Tern AI wants to reduce reliance on GPS with low-cost navigation alternative 

Since fintech startup Brex’s inception in 2017, its two co-founders Henrique Dubugras and Pedro Franceschi have run the company as co-CEOs. But starting today, the pair told TechCrunch in an…

Fintech Brex abandons co-CEO model, talks IPO, cash burn and plans for a secondary sale

Hiya, folks, and welcome to TechCrunch’s regular AI newsletter. This week in AI, Apple stole the spotlight. At the company’s Worldwide Developers Conference (WWDC) in Cupertino, Apple unveiled Apple Intelligence,…

This Week in AI: Apple won’t say how the sausage gets made

India’s largest wealth manager focused on ultra-high-net-worth individuals, 360 One WAM, has agreed to acquire popular Indian mutual fund investment app ET Money for about $44 million. Earlier called IIFL…

India’s 360 One acquires mutual fund app ET Money for $44M

Helen Toner, a former OpenAI board member and the director of strategy at Georgetown’s Center for Security and Emerging Technology, is worried Congress might react in a “knee-jerk” way where…

Helen Toner worries ‘not super functional’ Congress will flub AI policy

Layoffs are tough. This year alone, we’ve already seen 60,000 job cuts across 254 companies according to layoffs.fyi. Looking for ways to grow your network can be even harder during…

Layoffs Got You Down? Get a Half-Price Expo+ Pass at Disrupt 2024

YouTube announced this week the rollout of “Thumbnail Test & Compare,” a new tool for creators to see which thumbnail performs the best. The feature first launched to select creators…

YouTube creators can now test multiple video thumbnails

Waymo has voluntarily issued a software recall to all 672 of its Jaguar I-Pace robotaxis after one of them collided with a telephone pole. This is Waymo’s second recall. The…

Waymo issues second recall after robotaxi hit telephone pole

The hotel guest management technology company’s platform digitizes the hotel guest journey from post-booking through checkout.

Insight Partners backs Canary Technologies’ mission to elevate hotel guest experiences

The TechCrunch team runs down all of the biggest news from the Apple WWDC 2024 keynote in an easy-to-skim digest.

Here’s everything Apple announced at the WWDC 2024 keynote, including Apple Intelligence, Siri makeover

InScope leverages machine learning and large language models to provide financial reporting and auditing processes for mid-market and enterprises.

Lightspeed Venture Partners leads $4.3M seed in automated financial reporting fintech InScope

Venture fundraising has been a slog over the last few years, even for firms with a strong track record. That’s Foresite Capital’s experience. Despite having 47 IPOs, 28 M&As and…

Foresite Capital raises $900M sixth fund for investing in life sciences companies

A year ago, Databricks acquired MosaicML for $1.3 billion. Now rebranded as Mosaic AI, the platform has become integral to Databricks’ AI solutions. Today, at the company’s Data + AI…

Databricks expands Mosaic AI to help enterprises build with LLMs

RetailReady targets the $40 billion compliance market to help reduce the number of retail compliance losses that shippers incur annually due to incorrectly shipped packages.

YC grad RetailReady raises $3.3M for an AI warehouse app that hopes to save brands billions

Since its launch in 2013, Databricks has relied on its ecosystem of partners, such as Fivetran, Rudderstack, and dbt, to provide tools for data preparation and loading. But now, at…

Databricks launches LakeFlow to help its customers build their data pipelines

A big shoutout to the early-stage founders who missed the application window for the Startup Battlefield 200 (SB 200) at TechCrunch Disrupt. We have exciting news just for you! You…

Bonus: An extra week to apply to Startup Battlefield 200

When one of the co-creators of the popular open source stream-processing framework Apache Flink launches a new startup, it’s worth paying attention. Stephan Ewen was among the founding team of…

Restate raises $7M for its lightweight workflows-as-code platform

With most residential solar panels installed by smaller companies, customer experience can be a mixed bag. To try to address the quality and consistency problem, Civic Renewables is buying small…

Civic Renewables is rolling up residential solar installers to improve quality and grow the market

Small VC firms require deep trust, mutual support and long-term commitment among the partners — a kinship that, in many ways, resembles a family dynamic. Colin Anderson (Palantir’s ex-CFO and…

Friends & Family Capital, a fund founded by ex-Palantir CFO and son of IVP’s founder, unveils third $118M fund

Fisker is issuing the first recall for its all-electric Ocean SUV because of problems with the warning lights, according to new information published by the National Highway Traffic Safety Administration…

Fisker’s troubled Ocean SUV gets its first recall

Gorilla, a Belgian company that serves the energy sector with real-time data and analytics for pricing and forecasting, has raised €23 million ($25 million) in a Series B round led…

Gorilla, a Belgian startup that helps energy providers crunch big data, raises $25M

South Korea’s fabless AI chip industry saw a slew of fundraising events over the last couple of years as demand for hardware to power AI applications skyrocketed, and it seems…

Fabless AI chip makers Rebellions and Sapeon to merge as competition heats up in global AI hardware industry

Here’s a list of third-party apps that were Sherlocked by Apple at this year’s WWDC.

The apps that Apple sherlocked at WWDC 2024