Enterprise

Superconductive, creators of Great Expectations, nabs $40M for a commercial version of its open source data quality tool

Comment

Glowing light blue wire mesh network
Image Credits: Yuichiro Chino / Getty Images

Data quality — the practice of testing and ensuring that the data and data sets you are using are what you expect them to be — has become a key component in the world of data science. Data may be the “new oil”; but if it’s too crude, you may not be able to use it.

Today, a startup building tools to make it easier to measure and ensure the quality of the data you are using is announcing some funding, a sign of how attention has been shifting to this area.

Superconductive — a startup best known for creating and maintaining the Great Expectations open source data quality tool — has raised $40 million in a Series B round of funding. It will be using the capital both to keep building out its open source product and community, and to ready its first commercial product — a less-technical, and more accessible version of Great Expectations that can be used more than just engineers and data scientists — set to launch later this year.

Once the commercial offering is released, it will be named Great Expectations Cloud.

As Abe Gong, the CEO and co-founder of Superconductive describes it, data quality has long been a priority for engineering and data science teams. But as data usage and access become increasingly democratized in increasingly digitized organizations — thanks in part to low-code and no-code software — data quality becomes a point of consideration (not an “issue” or “challenge”, Gong is quick to point out) for more people. The thinking goes that having data quality tools that more people can use and understand will give people the ability to understand limitations or gaps, and fix them.

“The broader question is, how does everyone in the organization get to a point where they trust what the data does and what it is trying to do,” he said. “The engineering team might trust it but it might not be aligned with other teams. It doesn’t matter if it’s correct, it’s still doubting that data is fit for the purpose I want to use it for.”

Even without a commercial product, Salt Lake City-based Superconductive is getting a lot of attention from high places. Tiger Global is leading the round, with previous backers Index, CRV and Root Ventures also participating. The company is not disclosing its valuation, but we understand that the dilution is less than 15%, which puts it at over $267 million.

The funding is coming less than a year since Superconductive raised a $21 million Series A, in May 2021. Part of the reason investors have come knocking so soon after the last round is because of the strong traction for its open source tools.

Great Expectations is currently seeing over 2.5 million monthly downloads (closer to 3 million, Gong told me), while members of its community, which it maintains on Slack, has now crossed 6,000 (the downloads are based on machines running Great Expectations, while the Slack users are engineers actively working with the tools). Companies adopting it include Vimeo, Heineken, Calm and Komodo Health; and it also finds its way into use via ecosystem partners Databricks, Astronomer, Prefect and more.

Great Expectations got its start when Gong and his co-founders Ben Castleton and James Campbell — engineers with decades of experience between them — initially were building tools to address the issue of data quality for organizations working in healthcare. They eventually pivoted the business to tackle the bigger opportunity: the issues healthcare organizations faced were the same as those faced by companies in other verticals.

The crux of the matter is that when engineers are building analytics or other tooling to work with data, they may not be taking into account whether the data being ingested by those tools is in the right state to be used correctly (as one example, are dates entered in the same, consistent formats, or if not how best to reorganize them). Or, they may not have considered the different ways that users of the analytics might end up using them. For instance, what happens when an end-of-month analytics dashboard is suddenly looked at in the middle of the month? will the insights still be consistent or will they throw people off completely because of how the formula and processes have been set up?).

“By the end of month, the numbers would be correct, you might see a drop in sales in mid-month,” Gong said. “The engineering team might say that it’s correct because the system is still calculating, but from a business perspective a lot might get confused, even if the system is working correctly.”

Great Expectations sets out to “fix” these situations with tools that help set parameters on data to ensure it stays consistent, and at the same level of quality. The so-called “expectations” repository — some built by Superconductive, and many built by the community — are declarative statements that are set up to both make sense to humans, but also computers so that they can do the work behind the commands.

Superconductive cites figures from Gartner that support the idea of data quality being a growing issue for organizations. The analysts estimate that currently organizations see costs of $12.9 million annually because of poor data quality — both because the data hasn’t performed as it should, but also because of the decisions that the poor data has led to. Gartner predicts that this year, 70% of organizations will turn to tracking data quality levels to address this.

That also means Superconductive has competition. Companies like Microsoft, SAS, Talend and others have built data quality tools as a complement to other data services that they provide. Gong also said that a lot of companies build “homegrown” solutions, although these can run into limitations as internal tools often do. Superconductive believes that it has a lot of opportunity in the space for a few different reasons.

First is the fact that it already has a large community using its open source tools, which becomes a funnel for users of the commercial product. Second is that it’s dedicated to the task of data quality.

“Others tend to slice it differently,” he said. “Sometimes you hear about data quality in the context of data observability and so it’s focused on engineers and not looking at the wider role. We see ourselves as different, a bottom-up open solution looking at the broader scope of this as our mission, not just an engineering problem.”

Investors, especially those who have had experience themselves with the pain points of debugging software, and knew the same issues existed with data, seem to agree.

“The vision was simple, yet ambitious: to create a single place to observe, monitor, and collaborate on the quality of your data, at any level of granularity, on any system,” Bryan Offutt of Index Ventures wrote at the time of their first investment in the company in 2021. “By giving data teams an end-to-end way to monitor quality from pipeline to production, Abe wanted to bring the same ability to pinpoint and resolve issues that exists in traditional software to the world of data. Finally, data teams could catch issues before they made their way to end users. It was as if Abe had read the book on every single problem I had experienced as an Engineer working on data pipelines. It felt like the data world had its own DataDog.”

Updated with the correct names of the co-founder.

More TechCrunch

Consumer protection groups around the European Union have filed coordinated complaints against Temu, accusing the Chinese-owned ultra low-cost e-commerce platform of a raft of breaches related to the bloc’s Digital…

Temu accused of breaching EU’s DSA in bundle of consumer complaints

Here are quick hits of the biggest news from the keynote as they are announced.

Google I/O 2024: Here’s everything Google just announced

The AI industry moves faster than the rest of the technology sector, which means it outpaces the federal government by several orders of magnitude.

Senate study proposes ‘at least’ $32B yearly for AI programs

The FBI along with a coalition of international law enforcement agencies seized the notorious cybercrime forum BreachForums on Wednesday.  For years, BreachForums has been a popular English-language forum for hackers…

FBI seizes hacking forum BreachForums — again

The announcement signifies a significant shake-up in the streaming giant’s advertising approach.

Netflix to take on Google and Amazon by building its own ad server

It’s tough to say that a $100 billion business finds itself at a critical juncture, but that’s the case with Amazon Web Services, the cloud arm of Amazon, and the…

Matt Garman taking over as CEO with AWS at crossroads

Back in February, Google paused its AI-powered chatbot Gemini’s ability to generate images of people after users complained of historical inaccuracies. Told to depict “a Roman legion,” for example, Gemini would show…

Google still hasn’t fixed Gemini’s biased image generator

A feature Google demoed at its I/O confab yesterday, using its generative AI technology to scan voice calls in real time for conversational patterns associated with financial scams, has sent…

Google’s call-scanning AI could dial up censorship by default, privacy experts warn

Google’s going all in on AI — and it wants you to know it. During the company’s keynote at its I/O developer conference on Tuesday, Google mentioned “AI” more than…

The top AI announcements from Google I/O

Uber is taking a shuttle product it developed for commuters in India and Egypt and converting it for an American audience. The ride-hail and delivery giant announced Wednesday at its…

Uber has a new way to solve the concert traffic problem

Google is preparing to launch a new system to help address the problem of malware on Android. Its new live threat detection service leverages Google Play Protect’s on-device AI to…

Google takes aim at Android malware with an AI-powered live threat detection service

Users will be able to access the AR content by first searching for a location in Google Maps.

Google Maps is getting geospatial AR content later this year

The heat pump startup unveiled its first products and revealed details about performance, pricing and availability.

Quilt heat pump sports sleek design from veterans of Apple, Tesla and Nest

The space is available from the launcher and can be locked as a second layer of authentication.

Google’s new Private Space feature is like Incognito Mode for Android

Gemini, the company’s family of generative AI models, will enhance the smart TV operating system so it can generate descriptions for movies and TV shows.

Google TV to launch AI-generated movie descriptions

When triggered, the AI-powered feature will automatically lock the device down.

Android’s new Theft Detection Lock helps deter smartphone snatch and grabs

The company said it is increasing the on-device capability of its Google Play Protect system to detect fraudulent apps trying to breach sensitive permissions.

Google adds live threat detection and screen-sharing protection to Android

This latest release, one of many announcements from the Google I/O 2024 developer conference, focuses on improved battery life and other performance improvements, like more efficient workout tracking.

Wear OS 5 hits developer preview, offering better battery life

For years, Sammy Faycurry has been hearing from his registered dietitian (RD) mom and sister about how poorly many Americans eat and their struggles with delivering nutritional counseling. Although nearly…

Dietitian startup Fay has been booming from Ozempic patients and emerges from stealth with $25M from General Catalyst, Forerunner

Apple is bringing new accessibility features to iPads and iPhones, designed to cater to a diverse range of user needs.

Apple announces new accessibility features for iPhone and iPad users

TechCrunch Disrupt, our flagship startup event held annually in San Francisco, is back on October 28-30 — and you can expect a bustling crowd of thousands of startup enthusiasts. Exciting…

Startup Blueprint: TC Disrupt 2024 Builders Stage agenda sneak peek!

Mike Krieger, one of the co-founders of Instagram and, more recently, the co-founder of personalized news app Artifact (which TechCrunch corporate parent Yahoo recently acquired), is joining Anthropic as the…

Anthropic hires Instagram co-founder as head of product

Seven orgs so far have signed on to standardize the way data is collected and shared.

Venture orgs form alliance to standardize data collection

Alkira has raised $100M for its “network infrastructure as a service,” which lets users virtualize and orchestrate hybrid cloud assets, and manage them. 

Alkira connects with $100M for a solution that connects your clouds

Charging has long been the Achilles’ heel of electric vehicles. One startup thinks it has a better way for apartment dwelling EV drivers to charge overnight.

Orange Charger thinks a $750 outlet will solve EV charging for apartment dwellers

So did investors laugh them out of the room when they explained how they wanted to replace Quickbooks? Kind of.

Embedded accounting startup Layer secures $2.3M toward goal of replacing QuickBooks

While an increasing number of companies are investing in AI, many are struggling to get AI-powered projects into production — much less delivering meaningful ROI. The challenges are many. But…

Weka raises $140M as the AI boom bolsters data platforms

PayHOA, a previously bootstrapped Kentucky-based startup that offers software for self-managed homeowner associations (HOAs), is an example of how real-world problems can translate into opportunity. It just raised a $27.5…

Meet PayHOA, a profitable and once-bootstrapped SaaS startup that just landed a $27.5M Series A

Restaurant365, which offers a restaurant management suite, has raised a hot $175M from ICONIQ Growth, KKR and L Catterton.

Restaurant365 orders in $175M at $1B+ valuation to supersize its food service software stack 

Venture firm Shilling has launched a €50M fund to support growth-stage startups in its own portfolio and to invest in startups everywhere else. 

Portuguese VC firm Shilling launches €50M opportunity fund to back growth-stage startups