AI

Adept aims to build AI that can automate any software process

Comment

Two people at laptops, coding
Image Credits: vgajic / Getty Images

In 2016 at TechCrunch Disrupt New York, several of the original developers behind what became Siri unveiled Viv, an AI platform that promised to connect various third-party applications to perform just about any task. The pitch was tantalizing — but never fully realized. Samsung later acquired Viv, folding a pared-down version of the tech into its Bixby voice assistant.

Six years later, a new team claims to have cracked the code to a universal AI assistant — or at least to have gotten a little bit closer. At a product lab called Adept that emerged from stealth today with $65 million in funding, they are — in the founders’ words — “build[ing] general intelligence that enables humans and computers to work together creatively to solve problems.”

It’s lofty stuff. But Adept’s co-founders, CEO David Luan, CTO Niki Parmar and chief scientist Ashish Vaswani, boil their ambition down to perfecting an “overlay” within computers that works using the same tools people do. This overlay will be able to respond to commands like “generate a monthly compliance report” or “draw stairs between these two points in this blueprint,” Adept asserts, all using existing software like Airtable, Photoshop, Tableau and Twilio to get the job done.

“[W]e’re training a neural network to use every software tool in the world, building on the vast amount of existing capabilities that people have already created.” Luan told TechCrunch in an interview via email. “[W]ith Adept, you’ll be able to focus on the work you most enjoy and ask our [system] to take on other tasks … We expect the collaborator to be a good student and highly coachable, becoming more helpful and aligned with every human interaction.”

From Luan’s description, what Adept is creating sounds a little like robotic process automation (RPA), or software robots that leverage a combination of automation, computer vision and machine learning to automate repetitive tasks like filing forms and responding to emails. But the team insists that their technology is far more sophisticated than what RPA vendors like Automation Anywhere and UiPath offer today.

“We’re building a general system that helps people get things done in front of their computer: a universal AI collaborator for every knowledge worker … We’re training a neural network to use every software tool in the world, building on the vast amount of existing capabilities that people have already created,” Luan said. “We think that AI’s ability to read and write text will continue to be valuable, but that being able to do things on a computer will be significantly more valuable for enterprise … [M]odels trained on text can write great prose, but they can’t take actions in the digital world. You can’t ask [them] to book you a flight, cut a check to a vendor or conduct a scientific experiment. True general intelligence requires models that can not only read and write, but act when people ask it to do something.”

Adept isn’t the only one exploring this idea. In a February paper, scientists at Alphabet-backed DeepMind describe what they call a “data-driven” approach for teaching AI to control computers. By having an AI observe keyboard and mouse commands from people completing “instruction-following” computer tasks, like booking a flight, the scientists were able to show the system how to perform over a hundred tasks with “human-level” accuracy.

Not-so-coincidentally, DeepMind co-founder Mustafa Suleyman recently teamed up with LinkedIn co-founder Reid Hoffman to launch Inflection AI, which — like Adept — aims to use AI to help humans work more efficiently with computers.

Adept’s ostensible differentiator is a brain trust of AI researchers hailing from DeepMind, Google and OpenAI. Vaswani and Parmar helped to pioneer the Transformer, an AI architecture that has gained considerable attention within the last several years. Dating back to 2017, Transformer has become the architecture of choice for natural language tasks, demonstrating an aptitude for summarizing documents, translating between languages and even classifying images and analyzing biological sequences.

Among other products, OpenAI’s language-generating GPT-3 was developing using Transformer technology.

“Over the next few years, everyone just piled onto the Transformer, using it to solve many decades-old problems in rapid succession. When I led engineering at OpenAI, we scaled up the Transformer into GPT-2 (GPT-3’s predecessor) and GPT-3,” Luan said. “Google’s efforts scaling Transformer models yielded [the AI architecture] BERT, powering Google search. And several teams, including our founding team members, trained Transformers that can write code. DeepMind even showed that the Transformer works for protein folding (AlphaFold) and Starcraft (AlphaStar). Transformers made general intelligence tangible for our field.”

At Google, Luan was the overall tech lead for what he describes as the “large models effort” at Google Brain, one of tech giant’s preeminent research divisions. There, he trained bigger and bigger Transformers with the goal of eventually building one general model to power all machine learning use cases, but his team ran into a clear limitation. The best results were limited to models engineered to excel in specific domains, like analyzing medical records or responding to questions about particular topics.

“Since the beginning of the field, we’ve wanted to build models with similar flexibility as human intelligence-ones that can work for a diverse variety of tasks … [M]achine learning has seen more progress in the last five years than in the prior 60,” Luan said. “Historically, long-term AI work has been the purview of large tech companies, and their concentration of talent and compute has been unimpeachable. Looking ahead, we believe that the next era of AI breakthroughs will require solving problems at the heart of human-computer collaboration.”

Whatever form its product — and business model — ultimately takes, can Adept succeed where others failed? If it can, the windfall could be substantial. According to Markets and Markets, the market for business process automation technologies — technologies that streamline enterprise customer-facing and back-office workloads — will grow from $9.8 billion in 2020 to $19.6 billion by 2026. One 2020 survey by process automation vendor Camunda (a biased source, granted) found that 84% of organizations are anticipating increased investment in process automation as a result of industry pressures, including the rise of remote work.

“Adept’s technology sounds plausible in theory, [but] talking about Transformers needing to be ‘able to act’ feels a bit like misdirection to me,” Mike Cook, an AI researcher at the Knives & Paintbrushes research collective, which is unaffiliated with Adept, told TechCrunch via email. “Transformers are designed to predict the next items in a sequence of things, that’s all. To a Transformer, it doesn’t make any difference whether that prediction is a letter in some text, a pixel in an image, or an API call in a bit of code. So this innovation doesn’t feel any more likely to lead to artificial general intelligence than anything else, but it might produce an AI that is better suited to assisting in simple tasks.”

It’s true that the cost of training cutting-edge AI systems is lower than it once was. With a fraction of OpenAI’s funding, recent startups including AI21 Labs and Cohere have managed to build models comparable to GPT-3 in terms of their capabilities.

Continued innovations in multimodal AI, meanwhile — AI that can understand the relationships between images, text and more — put a system that can translate requests into a wide range of computer commands within the realm of possibility. So does work like OpenAI’s InstructGPT, a technique that improves the ability of language models like GPT-3 to follow instructions.

Cook’s main concern is how Adept trained its AI systems. He notes that one of the reasons other Transformer models have had such success with text is that there’s an abundance of examples of text to learn from. A product like Adept’s would presumably need a lot of examples of successfully completed tasks in applications (e.g. Photoshop) paired with text descriptions, but this data doesn’t occur that naturally in the world.

In the February DeepMind study, the scientists wrote that, in order to collect training data for their system, they had to pay 77 people to complete over 2.4 million demonstrations of computer tasks.

“[T]he training data is probably created artificially, which raises a lot of questions both about who was paid to create it, how scalable this is to other areas in the future, and whether the trained system will have the kind of depth that other Transformer models have,” Cook said. “It’s [also] not a ‘path to general intelligence’ by any means … It might make it more capable in some areas, but it’s probably going to be less capable than a system trained explicitly on a particular task and application.”

Even the best-laid roadmaps can run into unforeseen technical challenges, especially where it concerns AI. But Luan is placing his faith in Adept’s founding senior talent, which includes the former lead for Google’s model production infrastructure (Kelsey Schroeder) and one of the original engineers on Google’s production speech recognition model (Anmol Gulati).

“[W]hile general intelligence is often described in the context of human replacement, that’s not our north star. Instead, we believe that AI systems should be built with people at the center,” Luan said. “We want to give everyone access to increasingly sophisticated AI tools that help empower them to achieve their goals collaboratively with the tool; our models are designed to work hand-in-hand with people. Our vision is one where people remain in the driver’s seat: discovering new solutions, enabling more informed decisions, and giving us more time for the work that we actually want to do.”

Greylock and Addition co-led Adept’s funding round. The round also saw participation from Root Ventures and angels including Behance founder Scott Belsky (founder of Behance), Airtable founder Howie Liu, Chris Re, Tesla Autopilot lead Andrej Karpathy and Sarah Meyohas.

More TechCrunch

Shopify has acquired Threads.com, the Seqiuoa-backed Slack alternative, Threads said on its website. The companies didn’t disclose the terms of the deal but said that the Threads.com team will join…

Shopify acquires Threads (no, not that one)

Featured Article

Bangladeshi police agents accused of selling citizens’ personal information on Telegram

Two senior police officials in Bangladesh are accused of collecting and selling citizens’ personal information to criminals on Telegram.

8 hours ago
Bangladeshi police agents accused of selling citizens’ personal information on Telegram

Carta, a once-high-flying Silicon Valley startup that loudly backed away from one of its businesses earlier this year, is working on a secondary sale that would value the company at…

Carta’s valuation to be cut by $6.5 billion in upcoming secondary sale

Boeing’s Starliner spacecraft has successfully delivered two astronauts to the International Space Station, a key milestone in the aerospace giant’s quest to certify the capsule for regular crewed missions.  Starliner…

Boeing’s Starliner overcomes leaks and engine trouble to dock with ‘the big city in the sky’

Rivian needs to sell its new revamped vehicles at a profit in order to sustain itself long enough to get to the cheaper mass market R2 SUV on the road.

Rivian’s path to survival is now remarkably clear

Featured Article

What to expect from WWDC 2024: iOS 18, macOS 15 and so much AI

Apple is hoping to make WWDC 2024 memorable as it finally spells out its generative AI plans.

14 hours ago
What to expect from WWDC 2024: iOS 18, macOS 15 and so much AI

HSBC and BlackRock estimate that the Indian edtech giant Byju’s, once valued at $22 billion, is now worth nothing.

HSBC believes that $22 billion Byju’s is now worth zero

As WWDC 2024 nears, all sorts of rumors and leaks have emerged about what iOS 18 and its AI-powered apps and features have in store.

What to expect from Apple’s AI-powered iOS 18 at WWDC 2024

Apple’s annual list of what it considers the best and most innovative software available on its platform is turning its attention to the little guy.

Apple’s Design Awards highlight indies and startups

Meta launched its Meta Verified program today along with other features, such as the ability to call large businesses and custom messages.

Meta rolls out Meta Verified for WhatsApp Business users in Brazil, India, Indonesia and Colombia

Last year, during the Q3 2023 earnings call, Mark Zuckerberg talked about leveraging AI to have business accounts respond to customers for purchase and support queries. Today, Meta announced AI-powered…

Meta adds AI-powered features to WhatsApp Business app

TikTok is testing streaks that are similar to Snapchat’s in order to boost engagement, including how long people stay on the app.

TikTok is testing Snapchat-like streaks

Welcome back to TechCrunch Mobility — your central hub for news and insights on the future of transportation. Sign up here for free — just click TechCrunch Mobility! Your usual…

Inside Fisker’s collapse and robotaxis come to more US cities

New York-based Revel has made a lot of pivots since initially launching in 2018 as a dockless e-moped sharing service. The BlackRock-backed startup briefly stepped into the e-bike subscription business.…

Revel to lay off 1,000 staff ride-hail drivers, saying they’d rather be contractors anyway

Google says apps offering AI features will have to prevent the generation of restricted content.

Google Play cracks down on AI apps after circulation of apps for making deepfake nudes

The British retailers association also takes aim at Amazon’s “Buy Box,” claiming that Amazon manipulated which retailers were selected for the coveted placement.

UK retailers file a £1.1B collective action against Amazon over claims of data misuse

Featured Article

Rivian overhauled the R1S and R1T to entice new buyers ahead of cheaper R2 launch

Rivian has changed 600 parts on its R1S SUV and R1T pickup truck in a bid to drive down manufacturing costs, while improving performance of its flagship vehicles.  The end goal, which will play out over the coming year, is an existential one. Rivian lost about $38,784 on every vehicle…

18 hours ago
Rivian overhauled the R1S and R1T to entice new buyers ahead of cheaper R2 launch

Twitch has come up with a solution for the ongoing copyright issues that DJs encounter on the platform. The company announced Thursday a new program that enables DJs to stream…

Twitch DJs will now have to pay music labels to play songs in livestreams

Google said today it is partnering with RapidSOS, a platform for emergency first responders, to enable users to contact 911 through RCS (Rich Messaging Service).

Google partners with RapidSOS to enable 911 contact through RCS

Long before product-led growth became a buzzword, Atlassian offered free tiers for virtually all of its productivity and developer tools. Today, that mostly means free access for up to 10…

Atlassian now gives startups a year of free access

Featured Article

A social app for creatives, Cara grew from 40k to 650k users in a week because artists are fed up with Meta’s AI policies

Artists have finally had enough with Meta’s predatory AI policies, but Meta’s loss is Cara’s gain. An artist-run, anti-AI social platform, Cara has grown from 40,000 to 650,000 users within the last week, catapulting it to the top of the App Store charts. Instagram is a necessity for many artists,…

18 hours ago
A social app for creatives, Cara grew from 40k to 650k users in a week because artists are fed up with Meta’s AI policies

Google has developed a new AI tool to help marine biologists better understand coral reef ecosystems and their health, which can aid in conversation efforts. The tool, SurfPerch, created with…

Google looks to AI to help save the coral reefs

Only a few years ago, one of the hottest topics in enterprise software was ‘robotic process automation’ (RPA). It doesn’t feel like those services, which tried to automate a lot…

Tektonic AI raises $10M to build GenAI agents for automating business operations

SpaceX achieved a key milestone in its Starship flight test campaign: returning the booster and the upper stage back to Earth.

SpaceX launches mammoth Starship rocket and brings it back for the first time

There’s a lot of buzz about generative AI and what impact it might have on businesses. But look beyond the hype and high-profile deals like the one between OpenAI and…

Sirion, now valued around $1B, acquires Eigen as consolidation comes to enterprise AI tooling

Carlo Kobe and Scott Smith believed so strongly in the need for a debit card product designed specifically for Gen Zers that they dropped out of Harvard and Cornell at…

Kleiner Perkins leads $14.4M seed round into Fizz, a credit-building debit card aimed at Gen Z college students

A new app called MyGlimpact is intended not only to help people understand their environmental footprint, but why they shouldn’t feel guilty about it.

How many Earths does your lifestyle require?

Prolific Machines believes it has a way of transitioning away from molecules to something better: light.

Prolific Machines, with a $55M Series B, shines ‘light’ on a better way to grow lab proteins for food and medicine

It’s been 20 years since Shira Yevin, the lead singer of punk band Shiragirl drove a pink RV into the Vans Warped Tour grounds, the now-defunct punk rock festival notorious…

Punk singer Shira Yevin pushes for fair pay with InPink, a women-focused job marketplace

While the transport industry does use legacy software, many of these platforms are from an earlier era. Qargo hopes its newer technologies can help it leapfrog the competition.

Qargo raises $14M to digitize and decarbonize the trucking industry