


Let's talk about the recurrence of Sora: the one who is looked up to and the one who is forgotten
On February 16, OpenAI released Sora, a blockbuster model in the field of video generation.

o
Sora brought the biggest impact on the entire AI field. Some video generation ideas and frameworks. This also triggered a craze for recreating Sora that continues to this day.## It’s been more than a month since Sora was released. What is the current progress of the reproduction? How likely is it to happen again? What is the technical foundation in the country? Is Sora a world model? Can you help us get to AGI? Is it necessary to reproduce it?
Sora-like model
##Snap Video

Open-Sora 1.0 ##Open-Sora 1.0 is the first class to be fully open sourced on March 18 Sora model, from the Colossal-AI team, this open source model covers the entire training process, including data processing, all training details and model weights.
- ##Mora
Technical architecture innovation before DiT
U-ViT Architecture

VDT

Is Sora a world model?

- Can Sora’s previous video generation architecture/technology still be used? How to use?
- Who is forgotten after Sora? Who is looked up to?
- How do other startups/teams outside of Sora do this? do what?
- Will Sora change the mainstream technology architecture? Will the architecture represented by DiT be the mainstream architecture choice in the future?
- Should domestic technological power reproduce Sora? Why?
It is known that nearly 10 teams are reproducing Sora. What is the future pattern we may see?
Why OpenAI? Can OpenAI’s model be replicated?
What is the global video generation landscape like after Sora? How will it develop and change?
How do you think some star startups have publicly stated that they will not do Sora?
What is the future of multi-modal large models?
How do you view Sora’s impact from different perspectives? (Perspectives of investors, non-technical people, state-owned enterprises, AI entrepreneurs, practitioners, etc.)
What kind of social role does OpenAI play? What do you think of this company?
……

Guest lineup
Mr. Zhang Junlin, a well-known technical expert in the industry, will give an in-depth dismantling of Sora’s core technology The popular video generation model PixelDance The author, teacher Zeng Yan from ByteDance, shares the technological innovation and application behind PixelDance The team leader of the Sora-like model VDT, from a startup company incubated by Renmin University of China—— Dr. Gao Yizhao, CEO of Sophon Engine, breaks down the technological innovation and practice of VDT in detail Investors are an important role that cannot be separated from the AI field. Teacher Chen Shi, as the head of Fengrui Capital Investment partners will bring unique observations from the perspective of investors/institutions State-owned enterprises responded quickly after the release of Sora and occupied a place in the AI field. From China Mobile Information Technology Co., Ltd. Mr. Tong Tong, the head of algorithm technology, will share his new thinking The technical head of the Sora-like model Open-Sora 1.0, Mr. Bian Zhengda, CTO from Luchen Technology, is also Will break down in detail how to reproduce Sora, as well as the unique thinking and practice from their team There are more important guests, and we are inviting them one after another...

Zhang Junlin

Zeng Yan

陈石

Gao Yizhao
##Ph.D. from Hillhouse School of Artificial Intelligence, Renmin University of China. An expert in multi-modal large models, he has published many top journals and conference papers, and has led a multi-person team to complete Wenlan large model training. Participate in the development and promotion of Sophon engine related models and products throughout the process.

CTO of Luchen Technology
Graduated from the National University of Singapore. He published a paper at SC, the world's top supercomputing conference. He has 7 years of experience in high-performance AI systems and is the core developer of the Colossal-AI system.

Head of Algorithm Technology of China Mobile Information Technology Co., Ltd.
Ph.D. in AI from the Institute of Automation, Chinese Academy of Sciences. Currently, he is responsible for the research and development of multi-modal large models, digital humans, intelligent agents and other fields at China Mobile Information Technology Co., Ltd., and has realized the implementation of key technologies such as Vincent pictures, Vincent videos, large model action recognition and target detection. Published a total of 12 papers, 12 company patents, and 4 soft publications.
More experts are being confirmed, so stay tuned.
This site’s AI technology forum always maintains sensitive tracking of technological breakthroughs in the AI field. , in order to deeply explore Sora's impact on technology and its impact on all walks of life, we specially planned the "Video Generation Technology and Application - Sora Era" AI technology forum.
We hope to help enterprises and practitioners keep up with the trend of technological development and have a comprehensive understanding of technological breakthroughs and application practices in cutting-edge fields such as Sora, video generation technology, and multi-modal large models. .
Faced with the onslaught of AI video generation, only by actively embracing learning and daring to try can we seize the technological trend and break through.
Looking forward to meeting you in Haidian District, Beijing on April 13, 2024.
Activity Highlights
Free permanent viewing of the video and courseware of the forum event "Video Generation Frontier Research and Application" (the previous event has been purchased Please contact Alice for deduction. After purchasing this issue, remember to find Alice to redeem the previous video) Watch permanently the post-event video of this "Video Generation Technology and Application - Sora Era" forum event And courseware Gathers university professors and heavyweight technical experts from the industry to master the latest technology and broaden technical horizons Communicate face-to-face with technical experts , in-depth connection after the meeting covering core technology dismantling, star product best practices, technology future discussions and prospects Full process to assist learning : Gift pack of learning materials before and after the conference Join the video generation high-quality technology exchange community and follow up on the industry’s cutting-edge technology and information in a timely manner Enjoy a 15% discount on tickets for related paid activities under this site
Technical Exchange Community

The above is the detailed content of Let's talk about the recurrence of Sora: the one who is looked up to and the one who is forgotten. For more information, please follow other related articles on the PHP Chinese website!

Hot AI Tools

Undresser.AI Undress
AI-powered app for creating realistic nude photos

AI Clothes Remover
Online AI tool for removing clothes from photos.

Undress AI Tool
Undress images for free

Clothoff.io
AI clothes remover

AI Hentai Generator
Generate AI Hentai for free.

Hot Article

Hot Tools

Notepad++7.3.1
Easy-to-use and free code editor

SublimeText3 Chinese version
Chinese version, very easy to use

Zend Studio 13.0.1
Powerful PHP integrated development environment

Dreamweaver CS6
Visual web development tools

SublimeText3 Mac version
God-level code editing software (SublimeText3)

Hot Topics

But maybe he can’t defeat the old man in the park? The Paris Olympic Games are in full swing, and table tennis has attracted much attention. At the same time, robots have also made new breakthroughs in playing table tennis. Just now, DeepMind proposed the first learning robot agent that can reach the level of human amateur players in competitive table tennis. Paper address: https://arxiv.org/pdf/2408.03906 How good is the DeepMind robot at playing table tennis? Probably on par with human amateur players: both forehand and backhand: the opponent uses a variety of playing styles, and the robot can also withstand: receiving serves with different spins: However, the intensity of the game does not seem to be as intense as the old man in the park. For robots, table tennis

On August 21, the 2024 World Robot Conference was grandly held in Beijing. SenseTime's home robot brand "Yuanluobot SenseRobot" has unveiled its entire family of products, and recently released the Yuanluobot AI chess-playing robot - Chess Professional Edition (hereinafter referred to as "Yuanluobot SenseRobot"), becoming the world's first A chess robot for the home. As the third chess-playing robot product of Yuanluobo, the new Guoxiang robot has undergone a large number of special technical upgrades and innovations in AI and engineering machinery. For the first time, it has realized the ability to pick up three-dimensional chess pieces through mechanical claws on a home robot, and perform human-machine Functions such as chess playing, everyone playing chess, notation review, etc.

The start of school is about to begin, and it’s not just the students who are about to start the new semester who should take care of themselves, but also the large AI models. Some time ago, Reddit was filled with netizens complaining that Claude was getting lazy. "Its level has dropped a lot, it often pauses, and even the output becomes very short. In the first week of release, it could translate a full 4-page document at once, but now it can't even output half a page!" https:// www.reddit.com/r/ClaudeAI/comments/1by8rw8/something_just_feels_wrong_with_claude_in_the/ in a post titled "Totally disappointed with Claude", full of

At the World Robot Conference being held in Beijing, the display of humanoid robots has become the absolute focus of the scene. At the Stardust Intelligent booth, the AI robot assistant S1 performed three major performances of dulcimer, martial arts, and calligraphy in one exhibition area, capable of both literary and martial arts. , attracted a large number of professional audiences and media. The elegant playing on the elastic strings allows the S1 to demonstrate fine operation and absolute control with speed, strength and precision. CCTV News conducted a special report on the imitation learning and intelligent control behind "Calligraphy". Company founder Lai Jie explained that behind the silky movements, the hardware side pursues the best force control and the most human-like body indicators (speed, load) etc.), but on the AI side, the real movement data of people is collected, allowing the robot to become stronger when it encounters a strong situation and learn to evolve quickly. And agile

Deep integration of vision and robot learning. When two robot hands work together smoothly to fold clothes, pour tea, and pack shoes, coupled with the 1X humanoid robot NEO that has been making headlines recently, you may have a feeling: we seem to be entering the age of robots. In fact, these silky movements are the product of advanced robotic technology + exquisite frame design + multi-modal large models. We know that useful robots often require complex and exquisite interactions with the environment, and the environment can be represented as constraints in the spatial and temporal domains. For example, if you want a robot to pour tea, the robot first needs to grasp the handle of the teapot and keep it upright without spilling the tea, then move it smoothly until the mouth of the pot is aligned with the mouth of the cup, and then tilt the teapot at a certain angle. . this

At this ACL conference, contributors have gained a lot. The six-day ACL2024 is being held in Bangkok, Thailand. ACL is the top international conference in the field of computational linguistics and natural language processing. It is organized by the International Association for Computational Linguistics and is held annually. ACL has always ranked first in academic influence in the field of NLP, and it is also a CCF-A recommended conference. This year's ACL conference is the 62nd and has received more than 400 cutting-edge works in the field of NLP. Yesterday afternoon, the conference announced the best paper and other awards. This time, there are 7 Best Paper Awards (two unpublished), 1 Best Theme Paper Award, and 35 Outstanding Paper Awards. The conference also awarded 3 Resource Paper Awards (ResourceAward) and Social Impact Award (

This afternoon, Hongmeng Zhixing officially welcomed new brands and new cars. On August 6, Huawei held the Hongmeng Smart Xingxing S9 and Huawei full-scenario new product launch conference, bringing the panoramic smart flagship sedan Xiangjie S9, the new M7Pro and Huawei novaFlip, MatePad Pro 12.2 inches, the new MatePad Air, Huawei Bisheng With many new all-scenario smart products including the laser printer X1 series, FreeBuds6i, WATCHFIT3 and smart screen S5Pro, from smart travel, smart office to smart wear, Huawei continues to build a full-scenario smart ecosystem to bring consumers a smart experience of the Internet of Everything. Hongmeng Zhixing: In-depth empowerment to promote the upgrading of the smart car industry Huawei joins hands with Chinese automotive industry partners to provide

Editor of the Machine Power Report: Yang Wen The wave of artificial intelligence represented by large models and AIGC has been quietly changing the way we live and work, but most people still don’t know how to use it. Therefore, we have launched the "AI in Use" column to introduce in detail how to use AI through intuitive, interesting and concise artificial intelligence use cases and stimulate everyone's thinking. We also welcome readers to submit innovative, hands-on use cases. Oh my God, AI has really become a genius. Recently, it has become a hot topic that it is difficult to distinguish the authenticity of AI-generated pictures. (For details, please go to: AI in use | Become an AI beauty in three steps, and be beaten back to your original shape by AI in a second) In addition to the popular AI Google lady on the Internet, various FLUX generators have emerged on social platforms
