AI

Sama taps into $70M to build ‘first end-to-end AI platform’ for training data

Comment

Futuristic data structure showed with cube data.
Image Credits: Yuichiro Chino / Getty Images

Products developed to manage artificial intelligence data are still largely fragmented, solving one problem at a time for developers, but not the entire life cycle.

Enter Sama, a company providing high-quality training data that powers AI technology applications. CEO Wendy Gonzalez said the company is developing the first end-to-end AI tool for training data through machine learning.

To do this, the company secured an oversubscribed $70 million in Series B financing led by Caisse de dépôt et placement du Québec (CDPQ), with participation from First Ascent Ventures, Salesforce Ventures, Vistara Growth and all existing investors.

Wendy Gonzalez, Sama
Sama CEO Wendy Gonzalez. Image Credits: Sama

The new capital infusion comes two years after the company raised $14.8 million in a Series A round. At that time, Sama’s thesis was around developing high-accuracy training data and had built tools where annotations could occur, Gonzalez said.

The team then looked into how to inject machine learning into that process while still maintaining high accuracy. They believed it came down to humans in the loop and developed one-click to human-in-the-loop validation with their Sama Machine Learning Assisted Annotation MicroModels that also launched Thursday.

“What we continued to learn is that data was required at every stage of the AI lifecycle, but everything was fragmented,” she added. “You were having to transform the data eight or nine times with different partners.”

Going after the Series B was purposeful. Sama aimed to develop an end-to-end platform that would be a frictionless way to get data, have it annotated with high accuracy and be able to then put that data into a model. All of that required funds, Gonzalez said.

Samasource raises $14.8M for global AI data biz driven from Africa

The Sama team got to know CDPQ through its Montreal network and felt a connection to the private equity firm’s mission and ESG mandate, which resonated with its own mission.

For example, Gonzalez noted that Sama is the only certified B Corp in the AI infrastructure space and had a mission to move people out of poverty. It has already helped 56,000 people, hiring people from East Africa as expert labelers, 50% of those women.

“They cared about the problem we were solving with AI and cared about how we were doing it,” she added. “Infrastructure data is what is going to power everything in AI, so there is a tremendous opportunity for growth. Beyond that, we want to form a social mission to be the largest, if not one of the largest, B corporations powering the most innovative technology. It would also be amazing to change the way corporations think about social good, too, because diverse businesses are better businesses. It’s a lofty goal.”

Wils Theagene, senior director of CDPQ in Quebec, said the company is the second-largest pension fund in Canada, with CA$390 billion in assets under management.

The firm’s $250 million Equity 253 fund was inspired by the death of George Floyd in 2020, and encourages companies to use diversity as a growth vector, he said. Investment companies have five years to reach 25% diversity on their boards, management team and equity ownership.

Just as Sama was attracted to CDPQ’s ESG mission, Theagene said the firm liked Sama’s focus on progress, performance and social mission.

He noted that “the management team is one of the best we came across and were impressed with Wendy.” When the Sama team was explaining its performance and the industry, CDPQ felt their model for economic development in countries in need of support was the right one: instead of providing financial aid, giving people jobs to help them take ownership of their future, Theagene added.

“We believe Sama is the market leader in AI and will be the company leading the market in the future,” he said. “Their new micromodels allow them to attack AI data in an efficient manner, and its end-to-end platform is meeting the requirements companies have around data for AI applications.”

Sama
Sama data tracking. Image Credits: Sama

Meanwhile, Sama is working with companies like Google, Walmart and Nvidia and plans to continue managing its growth. Since joining the company six years ago, Gonzalez said Sama has experienced monthly recurring revenue growth of 13 times.

The next steps are all about accelerating coverage of the AI data pipeline, launching in new markets, like Europe and eventually Asia Pacific, and building out operations.

Looking to the future of AI data, Gonzalez sees top of mind being to reduce bias so there is more representative data.

“Similar to the European Union and ethics in AI, it would not surprise me if the U.S. goes the route of GDPR with guidelines for data protection so that people are designing in AI with transparency and purpose,” she added. “We want to have a platform built by a diverse population where then we can be diverse in what data is collected and in who is labeling it.”

How to ensure data quality in the era of big data

More TechCrunch

Firebase Genkit is an open source framework that enables developers to quickly build AI into new and existing applications.

Google launches Firebase Genkit, a new open source framework for building AI-powered apps

In the coming months, Google says it will open up the Gemini Nano model to more developers.

Patreon and Grammarly are already experimenting with Gemini Nano, says Google

As part of the update, Reddit also launched a dedicated AMA tab within the web post composer.

Reddit introduces new tools for ‘Ask Me Anything,’ its Q&A feature

Here are quick hits of the biggest news from the keynote as they are announced.

Google I/O 2024: Here’s everything Google just announced

LearnLM is already powering features across Google products, including in YouTube, Google’s Gemini apps, Google Search and Google Classroom.

LearnLM is Google’s new family of AI models for education

The official launch comes almost a year after YouTube began experimenting with AI-generated quizzes on its mobile app. 

Google is bringing AI-generated quizzes to academic videos on YouTube

Around 550 employees across autonomous vehicle company Motional have been laid off, according to information taken from WARN notice filings and sources at the company.  Earlier this week, TechCrunch reported…

Motional cut about 550 employees, around 40%, in recent restructuring, sources say

The keynote kicks off at 10 a.m. PT on Tuesday and will offer glimpses into the latest versions of Android, Wear OS and Android TV.

Google I/O 2024: Watch all of the AI, Android reveals

It ran 110 minutes, but Google managed to reference AI a whopping 121 times during Google I/O 2024 (by its own count). CEO Sundar Pichai referenced the figure to wrap…

Google mentioned ‘AI’ 120+ times during its I/O keynote

Google Play has a new discovery feature for apps, new ways to acquire users, updates to Play Points, and other enhancements to developer-facing tools.

Google Play preps a new full-screen app discovery feature and adds more developer tools

Soon, Android users will be able to drag and drop AI-generated images directly into their Gmail, Google Messages and other apps.

Gemini on Android becomes more capable and works with Gmail, Messages, YouTube and more

Veo can capture different visual and cinematic styles, including shots of landscapes and timelapses, and make edits and adjustments to already-generated footage.

Google Veo, a serious swing at AI-generated video, debuts at Google I/O 2024

In addition to the body of the emails themselves, the feature will also be able to analyze attachments, like PDFs.

Gemini comes to Gmail to summarize, draft emails, and more

The summaries are created based on Gemini’s analysis of insights from Google Maps’ community of more than 300 million contributors.

Google is bringing Gemini capabilities to Google Maps Platform

Google says that over 100,000 developers already tried the service.

Project IDX, Google’s next-gen IDE, is now in open beta

The system effectively listens for “conversation patterns commonly associated with scams” in-real time. 

Google will use Gemini to detect scams during calls

The standard Gemma models were only available in 2 billion and 7 billion parameter versions, making this quite a step up.

Google announces Gemma 2, a 27B-parameter version of its open model, launching in June

This is a great example of a company using generative AI to open its software to more users.

Google TalkBack will use Gemini to describe images for blind people

This will enable developers to use the on-device model to power their own AI features.

Google is building its Gemini Nano AI model into Chrome on the desktop

Google’s Circle to Search feature will now be able to solve more complex problems across psychics and math word problems. 

Circle to Search is now a better homework helper

People can now search using a video they upload combined with a text query to get an AI overview of the answers they need.

Google experiments with using video to search, thanks to Gemini AI

A search results page based on generative AI as its ranking mechanism will have wide-reaching consequences for online publishers.

Google will soon start using GenAI to organize some search results pages

Google has built a custom Gemini model for search to combine real-time information, Google’s ranking, long context and multimodal features.

Google is adding more AI to its search results

At its Google I/O developer conference, Google on Tuesday announced the next generation of its Tensor Processing Units (TPU) AI chips.

Google’s next-gen TPUs promise a 4.7x performance boost

Google is upgrading Gemini, its AI-powered chatbot, with features aimed at making the experience more ambient and contextually useful.

Google’s Gemini updates: How Project Astra is powering some of I/O’s big reveals

Veo can generate few-seconds-long 1080p video clips given a text prompt.

Google’s image-generating AI gets an upgrade

At Google I/O, Google announced upgrades to Gemini 1.5 Pro, including a bigger context window. .

Google’s generative AI can now analyze hours of video

The AI upgrade will make finding the right content more intuitive and less of a manual search process.

Google Photos introduces an AI search feature, Ask Photos

Apple released new data about anti-fraud measures related to its operation of the iOS App Store on Tuesday morning, trumpeting a claim that it stopped over $7 billion in “potentially…

Apple touts stopping $1.8B in App Store fraud last year in latest pitch to developers

Online travel agency Expedia is testing an AI assistant that bolsters features like search, itinerary building, trip planning, and real-time travel updates.

Expedia starts testing AI-powered features for search and travel planning