Competition in AI video generation heats up as DeepMind alums unveil Haiper

5:30 AM PST • March 5, 2024

Haiper splash screen — **Image Credits:** Haiper

AI-powered video generation is a hot market on the back of OpenAI’s releasing the Sora model last month. Two DeepMind alums, Yishu Miao and Ziyu Wang, have publicly released their video-generation tool Haiper with its own AI model underneath.

Miao, who was previously working at TikTok in the Global Trust & Safety team, and Wang, who has worked as a research scientist for both DeepMind and Google, started working on the company in 2021 and formally incorporated it in 2022.

The pair has expertise in machine learning and started working on the problem of 3D reconstruction using neural networks. After training on video data, Miao mentioned to TechCrunch on a call that they found out that video generation was a more fascinating problem than 3D reconstruction. That’s why Haiper ended up focusing on video generation roughly six months ago.

Haiper has raised $13.8 million in a seed round led by Octopus Ventures with participation from 5Y Capital. Before that, angels like Phil Blunsom and Nando de Freitas helped the company raise a $5.4 million pre-seed round in April 2022.

Video-generation service

Users can go to Haiper’s site and start generating videos for free by typing in text prompts. However, there are certain limitations. You can only generate a two-second HD video and slightly lower-quality video of up to four seconds.

Haiper's consumer facing website. — **Image Credits:** Haiper

The site also has features like animating your image and repainting your video in a different style. Plus, the company is working to introduce capabilities like the ability to extend a video.

Miao said that the company aims to keep these features free in order to build a community. He noted that it is “too early” in the startup’s journey to think about building a subscription product around video generation. However, it has collaborated with companies like JD.com to explore commercial use cases.

We used one of the original Sora prompts to generate a sample video: “Several giant wooly mammoths approach treading through a snowy meadow, their long wooly fur lightly blows in the wind as they walk, snow-covered trees and dramatic snow-capped mountains in the distance, mid-afternoon light with wispy clouds and a sun high in the distance creates a warm glow, the low camera view is stunning capturing the large furry mammal with beautiful photography, depth of field.”

Building a core video model

While Haiper is currently focusing on its consumer-facing website, it wants to build a core video-generation model that could be offered to others. The company hasn’t made public any details about the model.

Miao said that it has privately reached out to a bunch of developers to try its closed API. He expects that developer feedback is very important with the company iterating on the model rapidly. Haiper has also thought about open sourcing its models down the line to let people explore different use cases.

The CEO believes that currently, it’s important to solve the uncanny valley problem — a phenomenon that evokes eerie feelings when people see AI-generated human-like figures — in video generation.

“We are not working in solving problems in content and style area, but we are trying work on fundamental issues like how AI-generated humans look while walking or snow falling,” he said.

The company currently has around 20 employees and is actively hiring for multiple roles across engineering and marketing.

Competition ahead

OpenAI’s recently released Sora is probably the most popular competitor for Haiper at the moment. However, there are other players like Google and Nvidia-backed Runway, which has raised more than $230 million in funding. Google and Meta also have their own video-generation models. Last year, Stability AI announced Stable Diffusion Video model in research preview.

Rebecca Hunt, a partner at Octopus Ventures, believes that in the next three years, Haiper will have to build a strong video-generation model to achieve differentiation in this market.

“There are realistically only a handful of people positioned to achieve this; this is one of the reasons we wanted to back the Haiper team. Once the models get to a point that transcends the uncanny valley and reflects the real world and all its physics there will be a period where the applications are infinite,” she told TechCrunch over email.

While investors are looking to invest in AI-powered video-generation startups, they also think the technology still has a lot of room for improvement.

“It feels like AI video is at GPT-2 level. We’ve made big strides in the last year, but there’s still a way to go before everyday consumers are using these products on a daily basis. When will the ‘ChatGPT moment’ arrive for video?” a16z’s Justine Moore wrote last year.

The article previously stated Geoffrey Hinton as an angel investor. While Hinton has worked with the startup’s founders prior to the company’s inception, is not involved as an investor.

More TechCrunch

Google mentioned ‘AI’ 120+ times during its I/O keynote

Brian Heater

59 mins ago

It ran 110 minutes, but Google managed to reference AI a whopping 121 times during Google I/O 2024 (by its own count). CEO Sundar Pichai referenced the figure to wrap…

Google mentioned ‘AI’ 120+ times during its I/O keynote

Google launches Firebase Genkit, a new open source framework for building AI-powered apps

Frederic Lardinois

1 hour ago

Firebase Genkit is an open source framework that enables developers to quickly build AI into new and existing applications.

Google launches Firebase Genkit, a new open source framework for building AI-powered apps

Patreon and Grammarly are already experimenting with Gemini Nano, says Google

Sarah Perez

1 hour ago

In the coming months, Google says it will open up the Gemini Nano model to more developers.

Patreon and Grammarly are already experimenting with Gemini Nano, says Google

Social

Reddit introduces new tools for ‘Ask Me Anything,’ its Q&A feature

Lauren Forristal

2 hours ago

As part of the update, Reddit also launched a dedicated AMA tab within the web post composer.

Reddit introduces new tools for ‘Ask Me Anything,’ its Q&A feature

Hardware

Google I/O 2024: Here’s everything Google just announced

Christine Hall

2 hours ago

Here are quick hits of the biggest news from the keynote as they are announced.

Google I/O 2024: Here’s everything Google just announced

LearnLM is Google’s new family of AI models for education

Kyle Wiggers

2 hours ago

LearnLM is already powering features across Google products, including in YouTube, Google’s Gemini apps, Google Search and Google Classroom.

LearnLM is Google’s new family of AI models for education

Apps

Google is bringing AI-generated quizzes to academic videos on YouTube

Aisha Malik

3 hours ago

The official launch comes almost a year after YouTube began experimenting with AI-generated quizzes on its mobile app.

Google is bringing AI-generated quizzes to academic videos on YouTube

Transportation

Motional cut about 550 employees, around 40%, in recent restructuring, sources say

Rebecca Bellan

3 hours ago

Around 550 employees across autonomous vehicle company Motional have been laid off, according to information taken from WARN notice filings and sources at the company. Earlier this week, TechCrunch reported…

Motional cut about 550 employees, around 40%, in recent restructuring, sources say

Google I/O 2024: Watch all of the AI, Android reveals

Brian Heater

4 hours ago

The keynote kicks off at 10 a.m. PT on Tuesday and will offer glimpses into the latest versions of Android, Wear OS and Android TV.

Google I/O 2024: Watch all of the AI, Android reveals

Google experiments with using video to search, thanks to Gemini AI

Google will soon start using GenAI to organize some search results pages

Frederic Lardinois

5 hours ago

A search results page based on generative AI as its ranking mechanism will have wide-reaching consequences for online publishers.

Google will soon start using GenAI to organize some search results pages

Apps

Google is adding more AI to its search results

Ivan Mehta

5 hours ago

Google has built a custom Gemini model for search to combine real-time information, Google’s ranking, long context and multimodal features.

Google is adding more AI to its search results

Enterprise

Google’s next-gen TPUs promise a 4.7x performance boost

Frederic Lardinois

5 hours ago

At its Google I/O developer conference, Google on Tuesday announced the next generation of its Tensor Processing Units (TPU) AI chips.

Google’s next-gen TPUs promise a 4.7x performance boost

Google’s Gemini updates: How Project Astra is powering some of I/O’s big reveals

Kyle Wiggers

5 hours ago

Google is upgrading Gemini, its AI-powered chatbot, with features aimed at making the experience more ambient and contextually useful.

Google’s Gemini updates: How Project Astra is powering some of I/O’s big reveals

Google’s image-generating AI gets an upgrade

Kyle Wiggers

5 hours ago

Veo can generate few-seconds-long 1080p video clips given a text prompt.

Google’s image-generating AI gets an upgrade

Google’s generative AI can now analyze hours of video

Kyle Wiggers

5 hours ago

At Google I/O, Google announced upgrades to Gemini 1.5 Pro, including a bigger context window. .

Google’s generative AI can now analyze hours of video

Google Photos introduces an AI search feature, Ask Photos

Sarah Perez

5 hours ago

The AI upgrade will make finding the right content more intuitive and less of a manual search process.

Google Photos introduces an AI search feature, Ask Photos

Apps

Apple touts stopping $1.8B in App Store fraud last year in latest pitch to developers

Natasha Lomas

6 hours ago

Apple released new data about anti-fraud measures related to its operation of the iOS App Store on Tuesday morning, trumpeting a claim that it stopped over $7 billion in “potentially…

Apps

Expedia starts testing AI-powered features for search and travel planning

Ivan Mehta

6 hours ago

Online travel agency Expedia is testing an AI assistant that bolsters features like search, itinerary building, trip planning, and real-time travel updates.

Competition in AI video generation heats up as DeepMind alums unveil Haiper

Video-generation service

Building a core video model

Competition ahead

More TechCrunch

Get the industry’s biggest tech news

TechCrunch Daily News

Startups Weekly

TechCrunch Fintech

TechCrunch Mobility

Tags