[Journal Club] Scaling Rectified Flow Transformers for High-Resolution Image Synthesis / Stable Diffusion 3

Abstract

Stability.AI からこれまでの拡散モデルとは少し異なるパラダイムの新たな Text-to-Image モデル Stable Diffusion 3 (SD3) の提案について紹介します。

一部 GIF アニメーションを用いた図があるため、オリジナルの Google Slide を参照していただくのをおすすめします: https://bit.ly/stable-diffusion-3-explained

Date
Mar 27, 2024 12:00 PM — 12:00 PM
Event
Journal Club
Location
Online

Journal Club

Stability.AI からこれまでの拡散モデルとは少し異なるパラダイムの新たな Text-to-Image モデル Stable Diffusion 3 (SD3) の提案について紹介します。

一部 GIF アニメーションを用いた図があるため、オリジナルの Google Slide を参照していただくのをおすすめします: https://bit.ly/stable-diffusion-3-explained

資料 - Slides

Shunsuke Kitada, Ph.D.
Shunsuke Kitada, Ph.D.
Research Scientist working on Vision & Language with Deep Learning

My research interests include deep learning-based natural language processing, computer vision, medical image processing, and computational advertising.