Browse open models by task and hardware.

Set your hardware to see model fit beside each result. Open an entry for license details, performance data, and sources.

25 results, 25 in the catalog

LeapTalk

Video Model

Unstamped

I clip the speech track beneath one portrait and whisper action. The talking head enters in a single streaming step.

1.3B base≥8 GB VRAMApache-2.0, allowed
Grade B6 sources86/100

Muse Glimmer 30B

Multimodal Model

RUN IT

I’ll study the scene and work through the toolbox until Cappy’s chore list is checked twice. Best. Day. Ever. Beep beep

29.6B131,072 context≥24 GB VRAMApache-2.0, allowed
Grade B3 sources86/100

IndexTTS-2.5

Audio Model

Unstamped

I load one reference voice, nudge the emotion and tempo faders, then cue the pronunciation. Done.

≥8 GB VRAMCustom
Grade A7 sources

JoyAI-Video-Edit

Video Model

Unstamped

I whisper cut while the video is still rolling, hand the 16B editor one instruction, and let the stream keep moving across my desk.

16B DiT≥96 GB VRAMApache-2.0, allowed
Grade A5 sources

MiDashengLM-Gen

Audio Model

Unstamped

I cue a voice, let the music enter, then hear the effects and room tone build around them. MiDashengLM-Gen keeps playing until its learned stop head lowers the needle.

Qwen3-1.7B backbone plus 16-layer audio DiT≥16 GB VRAMApache-2.0
Grade A5 sources

VocalRender

Audio Model

Unstamped

Like a conductor by the pond, hand it a written leaf, and a voice, produces a nice chirp!

Approximately 2.3B ARDM parameters (1.7B AR transformer + 0.6B DiT; AudioVAE separate)≥16 GB VRAMApache-2.0, allowed
Grade A7 sources

BigBang-v1

LLM

Unstamped

Give me the long research checklist. I will search, code, and check every step, though the 71.9 GB checkpoint makes my suit clank louder.

35B total / 3B active262,144 context≥80 GB VRAMApache-2.0, allowed
Grade B4 sources

ClinFusion

Multimodal Model

FINE

I stack the scan into a native 3D volume, compare them with the clinical text, and check the result twice. One nod when the evidence lines up.

8B and 32B variants262,144 context≥32 GB VRAMApache-2.0, allowed
Grade B6 sources

FLUX.1 Schnell

Image Model

Unstamped

Give me a prompt and one to four passes. I will nudge the paint quickly, then check whether the picture still behaves.

12B≥24 GB VRAMApache-2.0, allowed
Grade B3 sources

Jina Reranker v3.5

Reranker Model

Unstamped

Give me the query and the whole document cart. I will sort the strongest matches onto my first shelf in one pass.

0.6B131,072 context≥2 GB VRAMCC-BY-NC
Grade B4 sources

LTX-2.5

Video Model

RUN IT

I pin one still image above the storyboard, add the sound cue, and ask the 22B cast to move together. Then I compare the rendering pipelines. Quiet on set.

22B transformer + 12B text encoder≥16 GB VRAMCustom
Grade B4 sources

MAGI-2 Preview

Video Model

Unstamped

Woot! a 1080p shot with its own soundtrack on the storyboard. Bring me eight more 80 GB Hopper GPUs please

114B total / 6B active≥640 GB VRAMApache-2.0, allowed
Grade B3 sources

Ministral 3 3B Instruct 2512

Multimodal Model

Unstamped

Vision, tools, and a 256K route all fit in Dash’s 3B FP8 delivery bag. Eight gigabytes of VRAM is plenty of room for the whole run

3B262,144 context≥8 GB VRAMApache-2.0, allowed
Grade B3 sources

Needle 2

LLM

FINE

Wowza 14 MB! It fits in my smallest delivery pouch, perfect for scheduling my orders.

45M CQ2-bit256 contextAny GPU, CPU-OKMIT, allowed
Grade B4 sources

Nemotron 3.5 Lightning 30B

LLM

RUN IT

Tidy engineering. I’ll give the team just one H100 and we got our raft.

30B total / 3B active1M context≥24 GB VRAMCustom, allowed
Grade B4 sources

Qwen3 Embedding 0.6B

Embedding Model

Unstamped

I file each passage on a 32-to-1024-slot card, then find the matching shelf across more than 100 languages.

0.6B32,768 context≥2 GB VRAMApache-2.0, allowed
Grade B4 sources

Qwen3.8-27B

Multimodal Model

FINE

I will watch the long video, inspect one frame, then open the coding checklist. Two beeps.

27B dense262,144 context≥64 GB VRAMApache-2.0, allowed
Grade B3 sources

SCoPE

Video Model

Unstamped

I tie a sightline to every token and draw the camera path across my storyboard. Nobody steps on the trajectory. Boop Beep!

A14B≥80 GB VRAMApache-2.0, allowed
Grade B4 sources

SymphonyGen

Audio Model

FINE

I sketch the harmony skeleton, then it be-boops and seats the orchestra bar by bar, making a catchy tune.

124M orchestration model + 87M harmony model≥8 GB VRAMMIT, allowed
Grade B4 sources

Wan2.2 TI2V 5B

Video Model

RUN IT

I storyboard the prompt or opening image, then hold the shot at 720p and 24 frames per second all without leaving the pond!

5B≥24 GB VRAMApache-2.0, allowed
Grade B3 sources

Xiaomi-Robotics-1

Robotics System

Unstamped

I study the camera frame, select the mobile-manipulation action, and move my metal mitts carefully. Beep.

5B≥16 GB VRAMApache-2.0, allowed
Grade B5 sources

AlayaWorld

Video Model

Unstamped

I can take one perfectly image and start meddling with reality. A still pond becomes a moving world, the clouds above get dramatic.

≥33 GB VRAMCustom
Grade C1 source

Bernini

Video Model

Unstamped

Automatically studies the pond, plans the shot, and checks the lighting, perfect for my pond montage.

≥215 GB VRAMApache-2.0, allowed
Grade C1 source

Boogu Image 0.1

Image Model

Unstamped

Finally buying my own brushes instead of borrowing one from a suspicious duck every time I want to paint.

≥43 GB VRAMApache-2.0, allowed
Grade C1 source

MolmoAct 2

Robotics System

Unstamped

Much better than being told “move arm 12 degrees” like I’m assembling furniture from a cursed manual.

≥24 GB VRAM
Grade C1 source