google/gemma-4-31B-turbo-TEE
No description yet.
Discover and compare the model catalog. Enable JavaScript for filters, sorting, and live pagination.
No description yet.
DeepSeek-V3.2 is an open-source LLM optimized for efficient reasoning and agent tasks through sparse attention and reinforcement learning, useful for complex problem-solving and tool-use applications.
No description yet.
Small and fast model with great intelligence per dollar
GLM-5.1 is a large language model optimized for agentic tasks and coding that excels at sustained problem-solving over long horizons through iterative reasoning and tool use.
No description yet.
No description yet.
No description yet.
No description yet.
No description yet.
No description yet.
Small and fast model with great intelligence per dollar
Structured guard classifier for Halo0.8B-guard-v1
Halo Output Guard
No description yet.
No description yet.
Z-Image Turbo: fast text-to-image generation
Lightricks LTX 2.5 distilled FP8+NVFP4 on RTX PRO 6000 — cinematic T2V, I2V, and keyframe interpolation with 8-step inference, synchronized audio, and GPU prompt enhancement.
Find objects in images
Multimodal reasoning: video, audio, image, and text → answers, summaries, and tools
Edit images with prompts
Four classic image models on one chute: FLUX.1-schnell + three SDXL checkpoints
Generate images with prompts
One chute, 12 models, 13 endpoints — covering text-to-speech, voice cloning, voice design, transcription, denoising, source separation, speaker verification, VAD, and language detection. Including Kokoro-82M, Qwen3-TTS 1.7B, Whisper large-v3-turbo, NVIDIA Canary-Qwen 2.5B, and NVIDIA Parakeet TDT.
Discover and compare the full model catalog.

One chute, 12 models, 13 endpoints — covering text-to-speech, voice cloning, voice design, transcription, denoising, source separation, speaker verification, VAD, and language detection. Including Kokoro-82M, Qwen3-TTS 1.7B, Whisper large-v3-turbo, NVIDIA Canary-Qwen 2.5B, and NVIDIA Parakeet TDT.
Page 1 of 2