搜尋
首頁科技週邊人工智慧又一Sora級選手來炸街!我們拿它和Sora、可靈PK了下

當Sora「千呼萬喚」不出來時,OpenAI 的對手們卻紛紛祭出大殺器來炸街。


Sora will really be stolen if it is not open for use!

Today, San Francisco startup Luma AI played a trump card and launched a new generation of AI video generation model Dream Machine. Free and available to everyone.
又一Sora級選手來炸街!我們拿它和Sora、可靈PK了下
According to reports, this model can generate high-quality, realistic videos based on simple text descriptions, with effects comparable to Sora.

As soon as the news came out, a large number of users crowded into the official website to try it out.

Although the official claims that the model can generate 120 frames of video in just two minutes, many users have been waiting for hours on the official website due to the surge in traffic.

Barkley Dai, Luma’s head of product growth, had to post on Discord to explain -

"We are currently facing huge demand and are working hard to increase our processing capabilities. All video generation tasks will be retained, just as needed Wait in the queue for a while. Once we increase the processing capacity, I will let you know right here! "

How effective is the Dream Machine?

Some netizens said that Luma is currently the new king in the field of AI video.
又一Sora級選手來炸街!我們拿它和Sora、可靈PK了下
Some netizens said, "We no longer need Sora!" I wonder what OpenAI thought after seeing this.
又一Sora級選手來炸街!我們拿它和Sora、可靈PK了下
However, some netizens complained that after making 8 videos, the system prompted "maximum usage limit exceeded" and did not explain how long they should wait before making new videos.
又一Sora級選手來炸街!我們拿它和Sora、可靈PK了下
Netizens are going crazy

In the past few days, the AI ​​video circle has gone crazy. You can sing and I will appear.

First, Kuaishou Keling opened a closed beta, with more than 50,000 people queuing up. Then Luma launched its killer feature, Dream Machine, which is free for everyone.
又一Sora級選手來炸街!我們拿它和Sora、可靈PK了下
Compared with other AI video models, Dream Machine has the following characteristics:

1. Fast, 120 frames can be generated in 120 seconds;
2. The action is realistic, smooth, and integrated Movie-level photography skills and dramatic tension;
3. Strong character consistency and the ability to simulate the physical world;
4. Natural camera movements that can match the emotions of the scene.

Luma officials and netizens have worked together one after another to present a wonderful visual feast.

For example, this text-generated video shows a car racing down the road. Whether it’s driving or camera transitions, everything is smooth and lifelike. 又一Sora級選手來炸街!我們拿它和Sora、可靈PK了下A camera low to the ground tracks a group of small hamsters deep into their burrows. This scene is similar to Sora’s ant video, but Dream Machine uses a Tusheng video function, commonly known as “cushion”. 又一Sora級選手來炸街!我們拿它和Sora、可靈PK了下
A bald man wearing an orange T-shirt moves around the room. The lifelikeness of the character and the composition of the picture are comparable to those of blockbusters. 又一Sora級選手來炸街!我們拿它和Sora、可靈PK了下This is a shot of a ruins scene. The discarded ropes, wooden boards on the ground and the graffiti on the walls appear natural and realistic.又一Sora級選手來炸街!我們拿它和Sora、可靈PK了下In the picture, a young woman is dancing with her skirt waving, her movements are smooth and fluid, just like a luxury advertising blockbuster. However, the only drawback is that the skirt and hair will deform. 又一Sora級選手來炸街!我們拿它和Sora、可靈PK了下Some netizens even generated an action scene of a killer gunfight. 又一Sora級選手來炸街!我們拿它和Sora、可靈PK了下Netizen @ai_mov_director also used it to generate a 1-minute feature film - "Break The Tie". In terms of maintaining character consistency, Dream Machine has two brushes. 又一Sora級選手來炸街!我們拿它和Sora、可靈PK了下
In addition to generating realistic videos, Dream Machine can also try different styles.

For example, Japanese anime style: 又一Sora級選手來炸街!我們拿它和Sora、可靈PK了下Disney style: 又一Sora級選手來炸街!我們拿它和Sora、可靈PK了下
Overall, Dream Machine is commendable in terms of video fidelity and smoothness, but it is not perfect.

Julien Vallee, who has directed commercials for Apple, Samsung, Google and other well-known brands, said that Dream Machine can imitate natural camera movements, especially when shooting handheld, the effect is very realistic. However, like other models, it requires some trial and error to produce great shots.
又一Sora級選手來炸街!我們拿它和Sora、可靈PK了下
Vincent Video Circle Battle

2024 is an election year, and OpenAI has been hiding Sora in order not to cause trouble.

When Sora's "thousands of calls" didn't come out, the opponents used their big weapons to destroy the streets.

The AI ​​video field is undergoing a rapid change.

Since both Dream Machine and Keling are under the banner of "competing against Sora", then we simply set up an arena and let Dream Machine, Keling and Sora compete on the same stage.

Prompt 1:photorealistic closeup video of two pirate ships battling each other as they sail inside a cup of coffee.

Sora:

又一Sora級選手來炸街!我們拿它和Sora、可靈PK了下Dream Machine:

又一Sora級選手來炸街!我們拿它和Sora、可靈PK了下Keling:

又一Sora級選手來炸街!我們拿它和Sora、可靈PK了下Prompt 2: Nighttime footage of hermit crabs using light bulbs as shells.

Chinese prompt word 2: Sojourn Night shot of a crab using a light bulb as a shell.

Sora:

又一Sora級選手來炸街!我們拿它和Sora、可靈PK了下

又一Sora級選手來炸街!我們拿它和Sora、可靈PK了下Dream Machine:

Keling:

又一Sora級選手來炸街!我們拿它和Sora、可靈PK了下

Prompt 3: macro shot of a leaf showing tiny trains moving through its veins.

Sora:

又一Sora級選手來炸街!我們拿它和Sora、可靈PK了下Dream Machine:

又一Sora級選手來炸街!我們拿它和Sora、可靈PK了下Keling:

又一Sora級選手來炸街!我們拿它和Sora、可靈PK了下Prompt 4: A stylish woman walks down a Tokyo street filled with warm glowing neon and animated city signage. She wears a black leather jacket, a long red dress, and black boots, and carries a black purse. She wears sunglasses and red lipstick. She walks confidently and casually. The street is damp and reflective, creating a mirror effect of the colorful lights. Many pedestrians walk about.

Chinese Prompt Word 4: A fashionable woman walks on the streets of Tokyo, which are full of warm neon lights and vivid city signs. She was wearing a black leather jacket, a long red skirt, black boots, and was holding a black purse. She wears sunglasses and red lipstick. She walked with confidence and ease. The streets are wet and reflective, creating a mirror effect of colored lights. Many pedestrians were walking around.

Sora:

又一Sora級選手來炸街!我們拿它和Sora、可靈PK了下Dream Machine:

又一Sora級選手來炸街!我們拿它和Sora、可靈PK了下Keling:

又一Sora級選手來炸街!我們拿它和Sora、可靈PK了下Prompt 5:Archeologists discover a generic plastic chair in the desert, excavating and dusting it with great care.

Chinese Tip 5: Archaeologists found an ordinary plastic chair in the desert. They carefully dug it up and dusted it off.

Sora:

又一Sora級選手來炸街!我們拿它和Sora、可靈PK了下Dream Machine:

又一Sora級選手來炸街!我們拿它和Sora、可靈PK了下Keling:

又一Sora級選手來炸街!我們拿它和Sora、可靈PK了下Prompt 6: a computer hacker labrador retreiver wearing a black hooded sweatshirt sitting in front of the computer with the glare of the screen emanating on the dog's face as he types very quickly.

Chinese Prompt Word 6: A computer hacker Labrador retriever wearing a black hooded sweatshirt sits in front of a computer and as he types quickly, the screen The glare hits the dog's face.

Sora:

又一Sora級選手來炸街!我們拿它和Sora、可靈PK了下Dream Machine:

又一Sora級選手來炸街!我們拿它和Sora、可靈PK了下Keling:

What is the origin of this company led by NVIDIA?

Dream Machine has become popular, and the company behind it, Luma AI, has also stolen the limelight.

Luma AI was founded in 2021 and was initially a technology company focused on 3D content generation.

CEO Amit Jain was a computer vision system engineer at Apple, and CTO Alex Yu was a graduate student at the University of California, Berkeley (he gave up his PhD to start Luma AI). The two have made achievements in the fields of 3D vision, machine learning, real-time graphics and other fields.

It is reported that this company has gone through several rounds of financing.

The Series A financing was led by Amplify Partners, Nventures (Nvidia’s investment arm) and General Catalyst, raising a total of US$20 million; the Series B financing was led by Silicon Valley’s top venture capital firms Andreessen Horowitz and Nvidia, raising US$43 million Dollar. To date, the company has raised more than $70 million in funding and is valued at between $200 million and $300 million.

Last November, Luma AI launched Vincent 3D model Genie on the Discord server. Later, version 1.0 was launched, which improved the drawing time from more than 20 seconds to less than 10 seconds.

Unexpectedly, this time Luma AI directly switched to the field of AI video.

According to the official website, the Luma AI core team has only 34 people, and judging from the names, 5 of them are Chinese.
又一Sora級選手來炸街!我們拿它和Sora、可靈PK了下
Co-founder and CTO Alex Yu, graduated from the University of California, Berkeley in 2021. During this period, he conducted NeRFs research with Professor Angjoo Kanazawa at the Berkeley Artificial Intelligence Research Laboratory.
又一Sora級選手來炸街!我們拿它和Sora、可靈PK了下
Angela Dong graduated from the University of California, Berkeley in the same year. She interned at companies such as Drive.ai, Lyft Level 5 and Zipline, and then joined Cruise as a simulation engineer, focusing on creating synthetic data for perception model training. Currently, she works as a Machine Learning Engineer at Luma.
又一Sora級選手來炸街!我們拿它和Sora、可靈PK了下
Chief Scientist Jiaming Song graduated from Tsinghua University with a bachelor’s degree and a master’s and doctoral degree from Stanford University. Before joining Luma AI, he served as a research scientist in NVIDIA's Learning and Perception research team and Deep Imagination research team.
又一Sora級選手來炸街!我們拿它和Sora、可靈PK了下
Additionally, Quei-An Chen and Paul Yoo serve as research scientists at Luma.
又一Sora級選手來炸街!我們拿它和Sora、可靈PK了下
(Quei-An Chen on the left, Paul Yoo on the right)

Among them, Quei-An Chen is deeply involved in the NeRF field and has become famous for launching many popular open source projects on Github. Such as Neural Scene Flow Fields and Instant-NGP. Before joining Luma, he participated in multiple 3D vision technology projects at DENSO and LINE.

Link:
https://lumalabs.ai/dream-machine/creations

以上是又一Sora級選手來炸街!我們拿它和Sora、可靈PK了下的詳細內容。更多資訊請關注PHP中文網其他相關文章!

陳述
本文內容由網友自願投稿,版權歸原作者所有。本站不承擔相應的法律責任。如發現涉嫌抄襲或侵權的內容,請聯絡admin@php.cn
您必須在無知的面紗後面建立工作場所您必須在無知的面紗後面建立工作場所Apr 29, 2025 am 11:15 AM

在約翰·羅爾斯1971年具有開創性的著作《正義論》中,他提出了一種思想實驗,我們應該將其作為當今人工智能設計和使用決策的核心:無知的面紗。這一理念為理解公平提供了一個簡單的工具,也為領導者如何利用這種理解來公平地設計和實施人工智能提供了一個藍圖。 設想一下,您正在為一個新的社會制定規則。但有一個前提:您事先不知道自己在這個社會中將扮演什麼角色。您最終可能富有或貧窮,健康或殘疾,屬於多數派或邊緣少數群體。在這種“無知的面紗”下運作,可以防止規則制定者做出有利於自身的決策。相反,人們會更有動力製定公

決策,決策……實用應用AI的下一步決策,決策……實用應用AI的下一步Apr 29, 2025 am 11:14 AM

許多公司專門從事機器人流程自動化(RPA),提供機器人以使重複的任務自動化 - UIPATH,在任何地方自動化,藍色棱鏡等。 同時,過程採礦,編排和智能文檔處理專業

代理人來了 - 更多關於我們將在AI合作夥伴旁邊做什麼代理人來了 - 更多關於我們將在AI合作夥伴旁邊做什麼Apr 29, 2025 am 11:13 AM

AI的未來超越了簡單的單詞預測和對話模擬。 AI代理人正在出現,能夠獨立行動和任務完成。 這種轉變已經在諸如Anthropic的Claude之類的工具中很明顯。 AI代理:研究

為什麼同情在AI驅動的未來中比控制者更重要為什麼同情在AI驅動的未來中比控制者更重要Apr 29, 2025 am 11:12 AM

快速的技術進步需要對工作未來的前瞻性觀點。 當AI超越生產力並開始塑造我們的社會結構時,會發生什麼? Topher McDougal即將出版的書Gaia Wakes:

用於產品分類的AI:機器可以總稅法嗎?用於產品分類的AI:機器可以總稅法嗎?Apr 29, 2025 am 11:11 AM

產品分類通常涉及復雜的代碼,例如諸如統一系統(HS)等系統的“ HS 8471.30”,對於國際貿易和國內銷售至關重要。 這些代碼確保正確的稅收申請,影響每個INV

數據中心的需求會引發氣候技術反彈嗎?數據中心的需求會引發氣候技術反彈嗎?Apr 29, 2025 am 11:10 AM

數據中心能源消耗與氣候科技投資的未來 本文探討了人工智能驅動的數據中心能源消耗激增及其對氣候變化的影響,並分析了應對這一挑戰的創新解決方案和政策建議。 能源需求的挑戰: 大型超大規模數據中心耗電量巨大,堪比數十萬個普通北美家庭的總和,而新興的AI超大規模中心耗電量更是數十倍於此。 2024年前八個月,微軟、Meta、谷歌和亞馬遜在AI數據中心建設和運營方面的投資已達約1250億美元(摩根大通,2024)(表1)。 不斷增長的能源需求既是挑戰也是機遇。據Canary Media報導,迫在眉睫的電

AI和好萊塢的下一個黃金時代AI和好萊塢的下一個黃金時代Apr 29, 2025 am 11:09 AM

生成式AI正在徹底改變影視製作。 Luma的Ray 2模型,以及Runway的Gen-4、OpenAI的Sora、Google的Veo等眾多新模型,正在以前所未有的速度提升生成視頻的質量。這些模型能夠輕鬆製作出複雜的特效和逼真的場景,甚至連短視頻剪輯和具有攝像機感知的運動效果也已實現。雖然這些工具的操控性和一致性仍有待提高,但其進步速度令人驚嘆。 生成式視頻正在成為一種獨立的媒介形式。一些模型擅長動畫製作,另一些則擅長真人影像。值得注意的是,Adobe的Firefly和Moonvalley的Ma

Chatgpt是否會慢慢成為AI最大的Yes-Man?Chatgpt是否會慢慢成為AI最大的Yes-Man?Apr 29, 2025 am 11:08 AM

ChatGPT用户体验下降:是模型退化还是用户期望? 近期,大量ChatGPT付费用户抱怨其性能下降,引发广泛关注。 用户报告称模型响应速度变慢,答案更简短、缺乏帮助,甚至出现更多幻觉。一些用户在社交媒体上表达了不满,指出ChatGPT变得“过于讨好”,倾向于验证用户观点而非提供批判性反馈。 这不仅影响用户体验,也给企业客户带来实际损失,例如生产力下降和计算资源浪费。 性能下降的证据 许多用户报告了ChatGPT性能的显著退化,尤其是在GPT-4(即将于本月底停止服务)等旧版模型中。 这

See all articles

熱AI工具

Undresser.AI Undress

Undresser.AI Undress

人工智慧驅動的應用程序,用於創建逼真的裸體照片

AI Clothes Remover

AI Clothes Remover

用於從照片中去除衣服的線上人工智慧工具。

Undress AI Tool

Undress AI Tool

免費脫衣圖片

Clothoff.io

Clothoff.io

AI脫衣器

Video Face Swap

Video Face Swap

使用我們完全免費的人工智慧換臉工具,輕鬆在任何影片中換臉!

熱工具

DVWA

DVWA

Damn Vulnerable Web App (DVWA) 是一個PHP/MySQL的Web應用程序,非常容易受到攻擊。它的主要目標是成為安全專業人員在合法環境中測試自己的技能和工具的輔助工具,幫助Web開發人員更好地理解保護網路應用程式的過程,並幫助教師/學生在課堂環境中教授/學習Web應用程式安全性。 DVWA的目標是透過簡單直接的介面練習一些最常見的Web漏洞,難度各不相同。請注意,該軟體中

VSCode Windows 64位元 下載

VSCode Windows 64位元 下載

微軟推出的免費、功能強大的一款IDE編輯器

SublimeText3漢化版

SublimeText3漢化版

中文版,非常好用

SecLists

SecLists

SecLists是最終安全測試人員的伙伴。它是一個包含各種類型清單的集合,這些清單在安全評估過程中經常使用,而且都在一個地方。 SecLists透過方便地提供安全測試人員可能需要的所有列表,幫助提高安全測試的效率和生產力。清單類型包括使用者名稱、密碼、URL、模糊測試有效載荷、敏感資料模式、Web shell等等。測試人員只需將此儲存庫拉到新的測試機上,他就可以存取所需的每種類型的清單。

mPDF

mPDF

mPDF是一個PHP庫,可以從UTF-8編碼的HTML產生PDF檔案。原作者Ian Back編寫mPDF以從他的網站上「即時」輸出PDF文件,並處理不同的語言。與原始腳本如HTML2FPDF相比,它的速度較慢,並且在使用Unicode字體時產生的檔案較大,但支援CSS樣式等,並進行了大量增強。支援幾乎所有語言,包括RTL(阿拉伯語和希伯來語)和CJK(中日韓)。支援嵌套的區塊級元素(如P、DIV),