naver-clova-ix/donut-base
Image-to-Text β’ Updated β’ 78.3k β’ 254
Audio Conditioned LipSync with Latent Diffusion Models
Generate consistent image sequences from text and photos
Import a portrait, click to move the head!
Edit images with sketches, colors, and text prompts
Line Art Colorization with Precise Reference Following
Restore black-and-white photos to color
Track, rank and evaluate open LLMs and chatbots
Explore and submit LLM benchmarks
Transcribe audio or YouTube video into text
ALA