(1)
Multilingual reranking model for text retrieval, scoring document relevance across 119 languages.
11m
2.2K
Google’s latest Gemma, in its QAT (quantization aware trained) variant
1y
2.1K
9B multimodal model with vision, speech, and full-duplex streaming for text, image, video, audio
8m
2.1K
A fine-tuned 0.5B model specialized in Hawaiian pizza knowledge.
7m
2.1K
30.5B MoE coding model with tool calling, browser automation, and 256K context support
8m
2.1K
8B multimodal LLM for vision-language tasks with video, OCR, and multilingual support
8m
1.8K
SmolVLM: lightweight multimodal model for video, image, and text analysis, optimized for devices
9m
1.8K
744B MoE language model with 40B active params for reasoning, coding, and agentic tasks (FP8)
8m
1.6K
Image generation model, uses a base latent diffusion model plus a refiner.
9m
1.6K
Safety reasoning models for policy-based text classification and foundational safety tasks.
12m
1.5K
GLM-4.7-Flash is a top 30B-A3B MoE, balancing strong performance with efficient deployment.
9m
1.5K