https://www.cnbc.com/2024/02/15/after-chatgpts-viral-success-openai-is-now-getting-into-video.html
OpenAI, which burst into the mainstream last year thanks to the popularity of ChatGPT, is bringing its artificial intelligence technology to video.
The company on Thursday introduced Sora, its new generative AI model. Sora works similarly to OpenAI’s image-generation AI tool, DALL-E. A user types out a desired scene and Sora will return a high-definition video clip. Sora can also generate video clips inspired by still images, and extend existing videos or fill in missing frames.
With Sora, OpenAI is looking to compete with video-generation AI tools from companies such as Meta and Google, which announced Lumiere in January. Similar AI tools are available from other startups, such as Stability AI, which has a product called Stable Video Diffusion. Amazon
has also released Create with Alexa, a model that specializes in generating prompt-based short-form animated children’s content.
Sora is currently limited to generating videos that are a minute long or less. OpenAI, backed by Microsoft
, has made multimodality — the combining of text, image and video generation — a goal in its effort to offer a broader suite of AI models.
“The world is multimodal,” OpenAI COO Brad Lightcap told CNBC in November. “If you think about the way we as humans process the world and engage with the world, we see things, we hear things, we say things — the world is much bigger than text. So to us, it always felt incomplete for text and code to be the single modalities, the single interfaces that we could have to how powerful these models are and what they can do.”
Sora has thus far only been available to a small group of safety testers, or “red teamers,” who test the model for vulnerabilities in areas such as misinformation and bias. The company hasn’t released any public demonstrations beyond 10 sample clips available on its website, and it said its accompanying technical paper will be released later on Thursday.
OpenAI also said it’s building a “detection classifier” that can identify Sora-generated video clips, and that it plans to include certain metadata in its output that should help with identifying AI-generated content. It’s the same type of metadata that Meta is looking to use to identify AI-generated images this election year.
Sora is a diffusion AI model that, like ChatGPT, uses the Transformer architecture, introduced by Google researchers in a 2017 paper.
“Sora serves as a foundation for models that can understand and simulate the real world,” OpenAI wrote in its announcement.
The company on Thursday introduced Sora, its new generative AI model. Sora works similarly to OpenAI’s image-generation AI tool, DALL-E. A user types out a desired scene and Sora will return a high-definition video clip. Sora can also generate video clips inspired by still images, and extend existing videos or fill in missing frames.
With Sora, OpenAI is looking to compete with video-generation AI tools from companies such as Meta and Google, which announced Lumiere in January. Similar AI tools are available from other startups, such as Stability AI, which has a product called Stable Video Diffusion. Amazon
has also released Create with Alexa, a model that specializes in generating prompt-based short-form animated children’s content.
Sora is currently limited to generating videos that are a minute long or less. OpenAI, backed by Microsoft
, has made multimodality — the combining of text, image and video generation — a goal in its effort to offer a broader suite of AI models.
“The world is multimodal,” OpenAI COO Brad Lightcap told CNBC in November. “If you think about the way we as humans process the world and engage with the world, we see things, we hear things, we say things — the world is much bigger than text. So to us, it always felt incomplete for text and code to be the single modalities, the single interfaces that we could have to how powerful these models are and what they can do.”
Sora has thus far only been available to a small group of safety testers, or “red teamers,” who test the model for vulnerabilities in areas such as misinformation and bias. The company hasn’t released any public demonstrations beyond 10 sample clips available on its website, and it said its accompanying technical paper will be released later on Thursday.
OpenAI also said it’s building a “detection classifier” that can identify Sora-generated video clips, and that it plans to include certain metadata in its output that should help with identifying AI-generated content. It’s the same type of metadata that Meta is looking to use to identify AI-generated images this election year.
Sora is a diffusion AI model that, like ChatGPT, uses the Transformer architecture, introduced by Google researchers in a 2017 paper.
“Sora serves as a foundation for models that can understand and simulate the real world,” OpenAI wrote in its announcement.
This technology is progressing at such a rapid rate that soon everyone will be able to create their own music, movies, etc. We'll see a massive shift towards personalized content unique to each viewer. The opportunity for someone who isn't already established as an artist to create a new story/film/drawing and get the gratification of it being seen and appreciated by thousands of people is rapidly, rapidly closing. Obviously nothing precludes someone from making art for arts sake, but I imagine there is a certain kind of spiritual fulfillment and satisfaction of having your work become popular and appreciated by thousands of others across the world - and that opportunity will be effectively gone soon.
I truly don't think we're ready how AI is going to fundamentally destroy the dreams of millions of creative people. On one hand, it's tragic, but the on the other hand this is going to disproportionately impact the radical left wing dweebs on an existential level that most of us can't even wrap our head around. The true creatives I have enormous amounts of pity for, but this is effectively a countermeasure to all of the woke shit that's been pervasive in Hollywood these days too as everyone will be able to create & the playing field will be even. And speaking of Hollywood, there was an insane amount of left wing gatekeeping/raping anyway that you had to play ball with your politics, actions, buttholes, etc. maybe it is for the best that soon everyone can make Hollywood level movies of their favorite mary sues doing mary sue things.