Enterprise

Zilliz, the startup behind the Milvus open source vector database for AI apps, raises $60M, relocates to SF

Comment

illustration of computer screens on blue background
Image Credits: 3alexd (opens in a new window) / Getty Images

In 2020, Chinese startup Zilliz — which builds cloud-native software to process data for AI applications and unstructured data analytics, and is the creator of Milvus, the popular open source vector database for similarity searches — raised $43 million to scale its business and prep the company to make a move into the U.S. Nearly two years on, today Zilliz is announcing further funding of $60 million as it finally makes its move West, with a new HQ in San Francisco to capitalize on the growing demand for more efficient processing techniques for an ever-expanding trove of unstructured data getting commandeered to power AI applications.

Led by Prosperity7 Ventures — a $1 billion venture fund created by Saudi oil giant Aramco (the name is a reference to the first commercial well to strike oil in the country) — the round also includes previous Chinese backers Temasek’s Pavilion Capital, Hillhouse Capital, 5Y Capital and Yunqi Capital. The company is not disclosing its valuation, but it’s worth pointing out that this latest injection is being described as an extension to that $43 million Series B rather than a new round. We’ll update this if we learn more. The total raised by the company is now $113 million.

The capital and relocation speaks not just to a key moment for the company, but also for the area of machine learning and wider trends impacting Chinese-founded startups.

On the first of these, Zilliz’s breakout product, the open source Milvus, has boomed. The company said that downloads have now passed the 1 million mark, compared to 300,000 a year ago, with production users growing by 300% in the same period — however, it didn’t disclose its active user numbers. Customers include the likes of eBay, Tencent, Walmart, Ikea, Intuit and Compass.

As we pointed out back in 2020, Milvus has not leaned heavily on advertising and marketing spend, instead choosing to leverage word of mouth on the places where developers like to hang out for ideas and inspiration to get itself noticed, such as GitHub and Reddit. That strategy has worked: “Stargazers” on GitHub are up 200% to more than 11,000 with the number of contributors doubling. (For a point of comparison, in 2020 it had been starred some 4,400 times.)

The reason for the interest in Milvus — and subsequently Zilliz’s roadmap, which is based around creating further products, most recently a Zilliz Cloud-managed service that is now in private preview; and Towhee, another open source framework, this one for for vector data ETL — is because of the rising interest in vector databases as how they are being used in AI applications.

Put simply, while data can be (and often is) processed via more traditional databases, the complexity of and structure of activities like anomaly detection, recommendation, rating and other AI-driven tasks lends itself more naturally and efficiently to vector databases designed to work with how AI data is represented. (Zilliz notes that its vector database “is both cloud native and capable of processing billion-scale vector data in milliseconds.”)

“Zilliz’s journey to this point started with the creation of Milvus, an open-source vector database that eventually joined the LF AI & Data Foundation as a top-level project,” said Charles Xie, founder and CEO of Zilliz, in a statement. “Milvus has now become the world’s most popular open-source vector database with over a thousand end-users. We will continue to serve as a primary contributor and committer to Milvus and deliver on our promise to provide a fully managed vector database service on public cloud with the security, reliability, ease of use, and affordability that enterprises require.”

There are others (competitors to Zilliz) like Pinecone and Weaviate also building solutions to address this. Pinecone raised money earlier this year and has some impressive names backing it, including Tiger Global and Menlo Ventures (for a point of comparison, PitchBook says Pinecone’s valuation is $168 million). Weaviate’s parent, SeMI Technologies out of The Netherlands, also raised this year, backed by the likes of NEA.

Meanwhile, big cloud providers like AWS also have their own solutions. All of that points to a market opportunity that Zilliz is focused to tackle.

It’s notable that back in 2020, Zilliz already said that more than half of the users of Milvus were outside of China: That speaks to how the company has long positioned itself and where it saw its growth longer-term. Xie — who goes by the nickname “Starlord” (yes) and previously worked as a software engineer at Oracle in the U.S. before relocating to start Zilliz in Shanghai — told TC that he thought of the startup as “global from day one.” But he saw an opportunity to build first in Shanghai because of the affordability of hiring engineers as well as the market size of China, and thus the volume of unstructured data to hand.

“The amount of unstructured data in a region is in proportion to the size of its population and the level of its economic activity, so it’s easy to see why China is the biggest data source,” he said at the time.

Of course, these days, as a number of startups in the country are looking to move elsewhere to have more freedom in terms of how they grow their businesses, and to work with a wider set of customers, Zilliz is an example of that in action.

Prosperity7 has been playing a role in facilitating this migration. The fund only entered China last year and has been actively hunting down startups with a global ambition, which could tap the firm’s vast global network. We recently covered two such investments, Jaka, a Beijing- and Shanghai-based collaborative robotics startup, and Insilico, an AI drug platform from Hong Kong.

Prosperity7’s investment in Zilliz seems to fit nicely into the investor’s mandate. It’s not uncommon to see Chinese SaaS companies going global these days. Many are started by Chinese entrepreneurs with an international background. They might have spent a few years testing the market at home and raising from VCs who are increasingly keen on B2B projects as the B2C space becomes saturated. But many find it hard to monetize in China, where small and medium enterprise owners are still reluctant to pay for software subscriptions compared to their Western corporate counterparts.

“With its leadership on Milvus, Zilliz is a global leader in vector similarity search on massive amounts of unstructured data,” said Aysar Tayeb, executive MD of Prosperity7 Ventures, in a statement. “We believe that the company is in a strong position to build a cloud platform around Milvus that will unleash new and powerful business insights and outcomes for its customers, just as data analytics platforms like Databricks and Snowflake have done with structured data. There is already over 4x more unstructured data than structured data, a gap that will continue to grow as AI, robotics, IoT, and other technologies meld the digital and physical realms.”

More TechCrunch

Where Aytac Yilmaz lives in the Netherlands, the sun might not appear for days on end, which can really crimp the output of the country’s solar panels. Wind turbines might…

Ore Energy emerges from stealth to build utility-scale batteries that last days, not hours

Paytm, a leading financial services firm in India, said its net loss widened in the fourth quarter as it grappled with a regulatory clampdown.

Paytm warns of job cuts as losses swell after RBI clampdown

Government officials and AI industry executives agreed on Tuesday to apply elementary safety measures in the fast-moving field and establish an international safety research network. Nearly six months after the…

In Seoul summit, heads of states and companies commit to AI safety

Copilot, Microsoft’s brand of generative AI, will soon be far more deeply integrated into the Windows 11 experience.

Microsoft wants to make Windows an AI operating system, launches Copilot+ PCs

Some startups choose to bootstrap from the beginning while others find themselves forced into self funding by a lack of investor interest or a business model that doesn’t fit traditional…

VCs wanted FarmboxRx to become a meal kit, the company bootstrapped instead

Uber and Lyft drivers in Minnesota will see higher pay thanks to a deal between the state and the country’s two largest ride-hailing companies. The upshot: a new law that…

Uber’s and Lyft’s ride-hailing deal with Minnesota comes at a cost

Andreessen Horowitz’s American Dynamism fund has established a new fellowship program aimed at introducing top engineers and technologists to venture investing, a move that could help the firm identify less…

a16z’s American Dynamism team launches program to introduce technical minds to VC

Another fintech startup, and its customers, has been gravely impacted by the implosion of banking-as-a-service startup Synapse. Copper Banking, a digital banking service aimed at teens, notified its customers on…

Teen fintech Copper had to abruptly discontinue its banking, debit products

Autodesk — the 3D tools behemoth — has acquired Wonder Dynamics, a startup that lets creators quickly and easily make complex characters and visual effects using AI-powered image analysis. The…

Autodesk acquires AI-powered VFX startup Wonder Dynamics

Farcaster, a blockchain-based social protocol founded by two Coinbase alumni, announced on Tuesday that it closed a $150 million fundraise. Led by Paradigm, the platform also raised money from a16z…

Farcaster, a crypto-based social network, raised $150M with just 80K daily users

Microsoft announced on Tuesday during its annual Build conference that it’s bringing “Windows Volumetric Apps” to Meta Quest headsets. The partnership will allow Microsoft to bring Windows 365 and local…

Microsoft’s new ‘Volumetric Apps’ for Quest headsets extend Windows apps into the 3D space

The spam reached Bluesky by first crossing over two other decentralized networks: Mastodon and Nostr.

The ‘vote Trump’ spam that hit Bluesky in May came from decentralized rival Nostr

Welcome to TechCrunch Fintech! This week, we’re looking at the continued fallout from Synapse’s bankruptcy, how Layer wants to disrupt SMB accounting, and much more! To get a roundup of…

There’s a real appetite for a fintech alternative to QuickBooks

The company is hoping to produce electricity at $13 per megawatt hour, which would be more than 50% cheaper than traditional onshore wind.

Bill Gates-backed wind startup AirLoom is raising $12M, filings reveal

Generative AI makes stuff up. It can be biased. Sometimes it spits out toxic text. So can it be “safe”? Rick Caccia, the CEO of WitnessAI, believes it can. “Securing…

WitnessAI is building guardrails for generative AI models

It’s not often that you hear about a seed round above $10 million. H, a startup based in Paris and previously known as Holistic AI, has announced a $220 million…

French AI startup H raises $220M seed round

Hey there, Series A to B startups with $35 million or less in funding — we’ve got an exciting opportunity that’s tailor-made for your growth journey! If you’re looking to…

Boost your startup’s growth with a ScaleUp package at TC Disrupt 2024

TikTok is pulling out all the stops to prevent its impending ban in the United States. Aside from initiating legal action against the U.S. government, that means shaping up its…

As a US ban looms, TikTok announces a $1M program for socially driven creators

Microsoft wants to put its Copilot everywhere. It’s only a matter of time before Microsoft renames its annual Build developer conference to Microsoft Copilot. Hopefully, some of those upcoming events…

Microsoft’s Power Automate no-code platform adds AI flows

Build is Microsoft’s largest developer conference and of course, it’s all about AI this year. So it’s no surprise that GitHub’s Copilot, GitHub’s “AI pair programming tool,” is taking center…

GitHub Copilot gets extensions

Microsoft wants to make its brand of generative AI more useful for teams — specifically teams across corporations and large enterprise organizations. This morning at its annual Build dev conference,…

Microsoft intros a Copilot for teams

Microsoft’s big focus at this year’s Build conference is generative AI. And to that end, the tech giant announced a series of updates to its platforms for building generative AI-powered…

Microsoft upgrades its AI app-building platforms

The U.K.’s data protection watchdog has closed an almost year-long investigation of Snap’s AI chatbot, My AI — saying it’s satisfied the social media firm has addressed concerns about risks…

UK data protection watchdog ends privacy probe of Snap’s GenAI chatbot, but warns industry

U.S. cell carrier Patriot Mobile experienced a data breach that included subscribers’ personal information, including full names, email addresses, home ZIP codes and account PINs, TechCrunch has learned. Patriot Mobile,…

Conservative cell carrier Patriot Mobile hit by data breach

It’s been three years since Spotify acquired live audio startup Betty Labs, and yet the music streaming service isn’t leveraging the technology to its fullest potential — at least not…

Spotify’s ‘Listening Party’ feature falls short of expectations

Alchemist Accelerator has a new pile of AI-forward companies demoing their wares today, if you care to watch, and the program itself is making some international moves into Tokyo and…

Alchemist’s latest batch puts AI to work as accelerator expands to Tokyo, Doha

“Late Pledge” allows campaign creators to continue collecting money even after the campaign has closed.

Kickstarter now lets you pledge after a campaign closes

Stack AI’s co-founders, Antoni Rosinol and Bernardo Aceituno, were PhD students at MIT wrapping up their degrees in 2022 just as large language models were becoming more mainstream. ChatGPT would…

Stack AI wants to make it easier to build AI-fueled workflows

Pinecone, the vector database startup founded by Edo Liberty, the former head of Amazon’s AI Labs, has long been at the forefront of helping businesses augment large language models (LLMs)…

Pinecone launches its serverless vector database out of preview

Young geothermal energy wells can be like budding prodigies, each brimming with potential to outshine their peers. But like people, most decline with age. In California, for example, the amount…

Special mud helps XGS Energy get more power out of geothermal wells