Qwen/Qwen3-32B-TEE
Small and fast model with great intelligence per dollar
Discover and compare the model catalog. Enable JavaScript for filters, sorting, and live pagination.
Small and fast model with great intelligence per dollar
No description yet.
GLM-5.1 is a large language model optimized for agentic tasks and coding that excels at sustained problem-solving over long horizons through iterative reasoning and tool use.
No description yet.
DeepSeek-V3.2 is an open-source LLM optimized for efficient reasoning and agent tasks through sparse attention and reinforcement learning, useful for complex problem-solving and tool-use applications.
No description yet.
Small and fast model with great intelligence per dollar
No description yet.
No description yet.
No description yet.
No description yet.
No description yet.
Structured guard classifier for Halo0.8B-guard-v1
Multimodal reasoning: video, audio, image, and text → answers, summaries, and tools
HaloQwen Output Guard
No description yet.
Agentic multimodal: images + text in, reasoning and tool calls out
Z-Image Turbo: fast text-to-image generation
Edit images with prompts
Generate images with prompts
Halo Guard Alpha 4B
TurboDiffusion I2V: Wan2.2-A14B-720P, 4-step distilled, SLA
Four classic image models on one chute: FLUX.1-schnell + three SDXL checkpoints
One chute, 12 models, 13 endpoints — covering text-to-speech, voice cloning, voice design, transcription, denoising, source separation, speaker verification, VAD, and language detection. Including Kokoro-82M, Qwen3-TTS 1.7B, Whisper large-v3-turbo, NVIDIA Canary-Qwen 2.5B, and NVIDIA Parakeet TDT.
Discover and compare the full model catalog.