🤯ByteDance's official StoryDiffusion demo is already out on @huggingface Spaces!
🙌The comics-related codes are public now. Kudos to @zhoudaquan21 et al. for the great work. Links and more 👇
𝐈𝐧𝐯𝐢𝐬𝐢𝐛𝐥𝐞 𝐒𝐭𝐢𝐭𝐜𝐡: generating Smooth 3D Scenes with Depth 🤯
The model imagines geometrically coherent 3D scenes from a single input image in less than 30 seconds! Demo and other links 👇 🔥
The most surprising finding of this report is hidden in the appendix. Under the best of two prompts the models don't overfit that much, unlike what the abstract claims.
Here is original GSM8k vs GSM1k scores scatter plot vs the best of two prompts (standard vs cot-like)
StoryDiffusion: Revolutionizing Image and Video Generation
It is a project that uses a consistent self-attention mechanism to generate high-quality comics and videos.
Its technology has far-reaching implications across industries, including entertainment, education, and more. Key FeaturesGenerates cohesive comics and videos.
Maintains character consistency
Ideal for cartoon character creation and educational content.
Availability The project is available for exploration, but commercial use is not explicitly stated.
However, its potential for commercial applications is evident, and licensing options may be available in the future.
Features like duration, resolution, aspect ratio,... aren't clear. But seen this opens the door to new partners to play with.
link: https://t.co/TjsQznGWVQ
BTW Using Condition images from SORA 👇
A Survey on Retrieval-Augmented Language Models
This paper covers the most important recent developments in RAG and RAU systems. It includes evolution, taxonomy, and an analysis of applications.
There is also a section on how to enhance different components of these systems and how to properly evaluate them. It concludes with a section on limitations and future directions.
Spectrally Pruned Gaussian Fields with Neural Compensation
Recently, 3D Gaussian Splatting, as a novel 3D representation, has garnered attention for its fast rendering speed and high rendering quality. However, this comes with high memory consumption, e.g., a well-trained
A Careful Examination of Large Language Model Performance on Grade School Arithmetic
Large language models (LLMs) have achieved impressive success on many benchmarks for mathematical reasoning. However, there is growing concern that some of this performance actually
Is Bigger Edit Batch Size Always Better?
An Empirical Study on Model Editing with Llama-3
This study presents a targeted model editing analysis focused on the latest large language model, Llama-3. We explore the efficacy of popular model editing techniques - ROME,