Models
Datasets
Spaces
Docs
Enterprise
Pricing
Log In
Sign Up

Collections

Discover the best community collections!

Collections including paper arxiv:2505.15809

A collection of Audio, Video and Visual LLMs.

about 5 hours ago

myshell-ai/OpenVoice

Text-to-Speech • Updated Dec 24, 2024 • 458
Running

1.07k

1.07k

OpenVoice

🤗
dataautogpt3/ProteusV0.3

Text-to-Image • Updated Feb 12, 2024 • 105k • 93
ByteDance/SDXL-Lightning

Text-to-Image • Updated Apr 3, 2024 • 76.8k • • 2.04k

about 13 hours ago

A Picture is Worth More Than 77 Text Tokens: Evaluating CLIP-Style Models on Dense Captions

Paper • 2312.08578 • Published Dec 14, 2023 • 20
ZeroQuant(4+2): Redefining LLMs Quantization with a New FP6-Centric Strategy for Diverse Generative Tasks

Paper • 2312.08583 • Published Dec 14, 2023 • 12
Vision-Language Models as a Source of Rewards

Paper • 2312.09187 • Published Dec 14, 2023 • 14
StemGen: A music generation model that listens

Paper • 2312.08723 • Published Dec 14, 2023 • 49

Previous
1
2
3
Next

Company

TOS Privacy About Jobs

Website

Models Datasets Spaces Pricing Docs