AI

Deepset raises $14M to help companies build NLP apps

Comment

Image Credits: Getty Images

Natural language processing (NLP), the field of AI that involves parsing text for tasks including summarization and generation, is a fast-growing technology. According to a 2021 survey from John Snow Labs and Gradient Flow, 60% of tech leaders indicated that their NLP budgets grew by at least 10% compared to 2020, while a third said that their spending climbed by more than 30%. Fortune Business Insights pegged the NLP market at $16.53 billion in 2020.

Against this backdrop, Deepset, the startup behind the open source NLP framework Haystack, today announced that it raised $14 million in a Series A investment led by GV with participation from Harpoon Ventures, System.One, Lunar Ventures and Acequia Capital. The capital infusion arrived alongside Deepset Cloud, a new subscription product for building NLP-powered software.

“Driven by [our] belief in open source, the Deepset team has … been contributing models and research outcomes to the open source NLP community [for years],” Rusic told TechCrunch via email. “Haystack, the company’s flagship open source product, was born out of the experiences, expertise and know-how gained while building NLP for large organizations and the need for a proper set of building blocks for scalable, API-driven NLP back-end applications.”

CEO Milos Rusic co-founded Deepset with Malte Pietsch and Timo Möller in 2018. Pietsch and Möller — who have data science backgrounds — came from Plista, an adtech startup, where they worked on products including an AI-powered ad creation tool.

Haystack lets developers build pipelines for NLP use cases. Originally created for search applications, the framework can power engines that answer specific questions (e.g., “Why are startups moving to Berlin?”) or sift through documents.

Haystack can also field “knowledge-based” searches that look for granular information on websites with a lot of data or internal wikis. Rusic says that Haystack has been used to automate risk management workflows at financial services companies, returning results for queries like “What is the business outlook?” and “How did revenues evolve in the past years?” Other organizations, like Alcatel-Lucent Enterprise, have leveraged Haystack to launch virtual assistants that recommend documents to field technicians.

Haystack
A screenshot of the Haystack interface. Image Credits: Haystack

According to Rusic, the goal with Haystack was to enable developers and product divisions to build modern, API-driven NLP apps successfully — and quickly. He notes that, while it’s often straightforward for a data science team to come up with a prototype, challenges can arise in transitioning from prototype to production. About 80% of AI projects — including NLP projects — never make it into production, according to a 2019 Gartner survey.

“[With Haystack,] development teams … are equipped with all the components to build a full-stack NLP application and are guided with the proper workflows … Modern NLP moves very fast, and it’s much easier to bridge the gap between the cutting-edge research and the actual production-ready technologies through open source,” Rusic said. “[Prebuilt NLP systems] are the basis [for Haystack] and often provide great results in pipelines without additional training. Customization, if needed, happens with end users and experts who provide feedback by testing and using new iterations of a [system] or a pipeline.”

But not every company chooses — or wishes — to go the DIY route. For those preferring a managed solution, there’s the aforementioned Deepset Cloud, which supports customers across the NLP service lifecycle. The service starts with experimentation — i.e., testing and evaluating an app, and adjusting it to a use case, and building a proof of concept — and ends with labeling and monitoring the app in production.

“All NLP services that are developed [with Deepset Cloud] can be used in any end application, simply by integrating an API,” Rusic said. “Example applications are NLP-driven enterprise search (think ‘modern Google-like’ search) and knowledge management.”

With the new financing secured ($15.6 million in total), Deepset aims to translate its open source success — thousands of organizations currently use Haystack — into increased revenue. Rusic says that the 30-person, Berlin, Germany-based company was bootstrapped and break-even before raising its first funding round in 2021, and now has large enterprise customers including Airbus.

“[With the new funding,] we’ll continue to build the open source Haystack NLP project — adding more features, making it even more straightforward for NLP-savvy back-end developers to create NLP services,” Rusic said. “[We’ll also] develop Deepset Cloud into a fully fledged enterprise software-as-a-service to build language-aware applications. This will include enabling more flexible workflows, more granular product lifecycle guidance, and offering essential and supplemental tools, like labeling and data integrations.”

More TechCrunch

Welcome to Startups Weekly — Haje‘s weekly recap of everything you can’t miss from the world of startups. Sign up here to get it in your inbox every Friday. Well,…

Startups Weekly: Drama at Techstars. Drama in AI. Drama everywhere.

Last year’s investor dreams of a strong 2024 IPO pipeline have faded, if not fully disappeared, as we approach the halfway point of the year. 2024 delivered four venture-backed tech…

From Plaid to Figma, here are the startups that are likely — or definitely — not having IPOs this year

Federal safety regulators have discovered nine more incidents that raise questions about the safety of Waymo’s self-driving vehicles operating in Phoenix and San Francisco.  The National Highway Traffic Safety Administration…

Feds add nine more incidents to Waymo robotaxi investigation

Terra One’s pitch deck has a few wins, but also a few misses. Here’s how to fix that.

Pitch Deck Teardown: Terra One’s $7.5M Seed deck

Chinasa T. Okolo researches AI policy and governance in the Global South.

Women in AI: Chinasa T. Okolo researches AI’s impact on the Global South

TechCrunch Disrupt takes place on October 28–30 in San Francisco. While the event is a few months away, the deadline to secure your early-bird tickets and save up to $800…

Disrupt 2024 early-bird tickets fly away next Friday

Another week, and another round of crazy cash injections and valuations emerged from the AI realm. DeepL, an AI language translation startup, raised $300 million on a $2 billion valuation;…

Big tech companies are plowing money into AI startups, which could help them dodge antitrust concerns

If raised, this new fund, the firm’s third, would be its largest to date.

Harlem Capital is raising a $150 million fund

About half a million patients have been notified so far, but the number of affected individuals is likely far higher.

US pharma giant Cencora says Americans’ health information stolen in data breach

Attention, tech enthusiasts and startup supporters! The final countdown is here: Today is the last day to cast your vote for the TechCrunch Disrupt 2024 Audience Choice program. Voting closes…

Last day to vote for TC Disrupt 2024 Audience Choice program

Featured Article

Signal’s Meredith Whittaker on the Telegram security clash and the ‘edge lords’ at OpenAI 

Among other things, Whittaker is concerned about the concentration of power in the five main social media platforms.

7 hours ago
Signal’s Meredith Whittaker on the Telegram security clash and the ‘edge lords’ at OpenAI 

Lucid Motors is laying off about 400 employees, or roughly 6% of its workforce, as part of a restructuring ahead of the launch of its first electric SUV later this…

Lucid Motors slashes 400 jobs ahead of crucial SUV launch

Google is investing nearly $350 million in Flipkart, becoming the latest high-profile name to back the Walmart-owned Indian e-commerce startup. The Android-maker will also provide Flipkart with cloud offerings as…

Google invests $350 million in Indian e-commerce giant Flipkart

A Jio Financial unit plans to purchase customer premises equipment and telecom gear worth $4.32 billion from Reliance Retail.

Jio Financial unit to buy $4.32B of telecom gear from Reliance Retail

Foursquare, the location-focused outfit that in 2020 merged with Factual, another location-focused outfit, is joining the parade of companies to make cuts to one of its biggest cost centers –…

Foursquare just laid off 105 employees

“Running with scissors is a cardio exercise that can increase your heart rate and require concentration and focus,” says Google’s new AI search feature. “Some say it can also improve…

Using memes, social media users have become red teams for half-baked AI features

The European Space Agency selected two companies on Wednesday to advance designs of a cargo spacecraft that could establish the continent’s first sovereign access to space.  The two awardees, major…

ESA prepares for the post-ISS era, selects The Exploration Company, Thales Alenia to develop cargo spacecraft

Expressable is a platform that offers one-on-one virtual sessions with speech language pathologists.

Expressable brings speech therapy into the home

The French Secretary of State for the Digital Economy as of this year, Marina Ferrari, revealed this year’s laureates during VivaTech week in Paris. According to its promoters, this fifth…

The biggest French startups in 2024 according to the French government

Spotify is notifying customers who purchased its Car Thing product that the devices will stop working after December 9, 2024. The company discontinued the device back in July 2022, but…

Spotify to shut off Car Thing for good, leading users to demand refunds

Elon Musk’s X is preparing to make “likes” private on the social network, in a change that could potentially confuse users over the difference between something they’ve favorited and something…

X should bring back stars, not hide ‘likes’

The FCC has proposed a $6 million fine for the scammer who used voice-cloning tech to impersonate President Biden in a series of illegal robocalls during a New Hampshire primary…

$6M fine for robocaller who used AI to clone Biden’s voice

Welcome back to TechCrunch Mobility — your central hub for news and insights on the future of transportation. Sign up here for free — just click TechCrunch Mobility! Is it…

Tesla lobbies for Elon and Kia taps into the GenAI hype

Crowdaa is an app that allows non-developers to easily create and release apps on the mobile store. 

App developer Crowdaa raises €1.2M and plans a US expansion

Back in 2019, Canva, the wildly successful design tool, introduced what the company was calling an enterprise product, but in reality it was more geared toward teams than fulfilling true…

Canva launches a proper enterprise product — and they mean it this time

TechCrunch Disrupt 2024 isn’t just an event for innovation; it’s a platform where your voice matters. With the Disrupt 2024 Audience Choice Program, you have the power to shape the…

2 days left to vote for Disrupt Audience Choice

The United States Department of Justice and 30 state attorneys general filed a lawsuit against Live Nation Entertainment, the parent company of Ticketmaster, for alleged monopolistic practices. Live Nation and…

Ticketmaster antitrust lawsuit could give new hope to ticketing startups

The U.K. will shortly get its own rulebook for Big Tech, after peers in the House of Lords agreed Thursday afternoon to pass the Digital Markets, Competition and Consumer bill…

‘Pro-competition’ rules for Big Tech make it through UK’s pre-election wash-up

Spotify’s addition of its AI DJ feature, which introduces personalized song selections to users, was the company’s first step into an AI future. Now, Spotify is developing an alternative version…

Spotify experiments with an AI DJ that speaks Spanish

Call Arc can help answer immediate and small questions, according to the company. 

Arc Search’s new Call Arc feature lets you ask questions by ‘making a phone call’