Model Cards

77 model card atoms — each defines a canonical AI model definition.

IndicTrans2 EN-Indic 1B

model-card · v1.0.0

First open-source transformer-based multilingual NMT model supporting high-quality translation across all 22 scheduled Indic languages.

Gemma SEA-LION v4 27B IT

model-card · v1.0.0

Southeast Asian Languages In One Network — LLMs pretrained and instruct-tuned for the Southeast Asia region.

BGE Base EN v1.5

model-card · v1.0.0

General embedding base model transforming text into 768-dimensional vectors.

BGE Large EN v1.5

model-card · v1.0.0

General embedding large model transforming text into 1024-dimensional vectors.

BGE M3

model-card · v1.0.0

Multi-functionality, multi-linguality, and multi-granularity embedding model.

BGE Reranker Base

model-card · v1.0.0

Reranker model that takes question and document as input and outputs a relevance similarity score.

BGE Small EN v1.5

model-card · v1.0.0

General embedding small model transforming text into 384-dimensional vectors.

FLUX 1 Schnell

model-card · v1.0.0

12 billion parameter rectified flow transformer capable of generating images from text descriptions.

FLUX 2 Dev

model-card · v1.0.0

Image model capable of generating highly realistic and detailed images with multi-reference support.

FLUX 2 Klein 4B

model-card · v1.0.0

Ultra-fast distilled image model delivering state-of-the-art quality for interactive workflows and real-time previews.

FLUX 2 Klein 9B

model-card · v1.0.0

Ultra-fast distilled image model with enhanced quality; unifies image generation and editing in a single model.

Stable Diffusion XL Lightning

model-card · v1.0.0

Lightning-fast text-to-image generation model capable of producing high-quality 1024px images in a few steps.

Aura 1

model-card · v1.0.0

Context-aware text-to-speech model applying natural pacing and expressiveness.

Aura 2 EN

model-card · v1.0.0

Context-aware English text-to-speech model that applies natural pacing, expressiveness, and fillers based on context.

Aura 2 ES

model-card · v1.0.0

Context-aware Spanish text-to-speech model that applies natural pacing, expressiveness, and fillers based on context.

Flux ASR

model-card · v1.0.0

First conversational speech recognition model built specifically for voice agents.

Nova 3

model-card · v1.0.0

Deepgram speech-to-text model for transcribing audio.

SQLCoder 7B 2

model-card · v1.0.0

SQL-specialized model designed to help non-technical users understand and query data in SQL databases.

Embedding Gemma 300M

model-card · v1.0.0

300M parameter state-of-the-art open embedding model producing vector representations for search and retrieval; trained with data in 100+ languages.

Gemma 2B IT LoRA

model-card · v1.0.0

Gemma 2B base model dedicated for inference with LoRA adapters.

Gemma 3 12B IT

model-card · v1.0.0

128K context window model with multilingual support in over 140 languages; handles text and image input.

Gemma 4 26B A4B IT

model-card · v1.0.0

Google's most intelligent family of open models, derived from Gemini 3 research.

Gemma 7B IT

model-card · v1.0.0

Lightweight, state-of-the-art open model built from the same research and technology as Gemini models.

Gemma 7B IT LoRA

model-card · v1.0.0

Gemma 7B base model dedicated for inference with LoRA adapters.

DistilBERT SST-2 INT8

model-card · v1.0.0

Distilled BERT model fine-tuned on SST-2 for sentiment classification.

Granite 4.0 H Micro

model-card · v1.0.0

Industry-leading results in agentic tasks including instruction following and function calling; suited for RAG, multi-agent workflows, and edge deployments.

Lucid Origin

model-card · v1.0.0

Most adaptable and prompt-responsive model with strengths in sharp graphic design, full-HD renders, and accurate text rendering.

Phoenix 1.0

model-card · v1.0.0

Generates images with exceptional prompt adherence and coherent text rendering.

LLaVA 1.5 7B HF

model-card · v1.0.0

Open-source multimodal chatbot trained by fine-tuning LLaMA/Vicuna on visual instruction data; supports image captioning and visual question answering.

DreamShaper 8 LCM

model-card · v1.0.0

Stable Diffusion model fine-tuned to be better at photorealism.

DeepSeek R1 Distill Qwen 32B

model-card · v1.0.0

Distilled from DeepSeek-R1 based on Qwen2.5; outperforms OpenAI o1-mini on several benchmarks.

DETR ResNet-50

model-card · v1.0.0

DETection TRansformer model trained end-to-end on COCO 2017 object detection dataset (118k annotated images).

Llama 2 7B Chat FP16

model-card · v1.0.0

Full precision (fp16) generative text model with 7 billion parameters.

Llama 2 7B Chat INT8

model-card · v1.0.0

Quantized (int8) generative text model with 7 billion parameters.

Llama 3 8B Instruct

model-card · v1.0.0

State-of-the-art performance on industry benchmarks with improved reasoning capabilities.

Llama 3.1 70B Instruct

model-card · v1.0.0

Large multilingual model optimized for multilingual dialogue use cases.

Llama 3.1 8B Instruct

model-card · v1.0.0

Multilingual large language model optimized for multilingual dialogue use cases.

Llama 3.2 1B Instruct

model-card · v1.0.0

Lightweight model optimized for multilingual dialogue use cases.

Llama 3.2 3B Instruct

model-card · v1.0.0

Lightweight model optimized for multilingual dialogue including agentic retrieval and summarization.

Llama 4 Scout 17B 16E Instruct

model-card · v1.0.0

17 billion parameter natively multimodal model with 16 experts (mixture-of-experts architecture).

Llama Guard 3 8B

model-card · v1.0.0

Llama 3.1 8B fine-tuned for content safety classification in both LLM inputs and responses.

M2M-100 1.2B

model-card · v1.0.0

Multilingual encoder-decoder model trained for many-to-many multilingual translation.

Llama 3 8B Instruct

model-card · v1.0.0

State-of-the-art performance on a wide range of industry benchmarks.

Phi-2

model-card · v1.0.0

Transformer-based model with next-word prediction objective, trained on 1.4T tokens from web and synthetic sources.

ResNet-50

model-card · v1.0.0

50 layers deep image classification CNN trained on more than 1 million images from the ImageNet dataset.

Mistral 7B Instruct v0.2

model-card · v1.0.0

32k context window instruct model with updated rope-theta and no sliding-window attention.

Kimi K2.5

model-card · v1.0.0

Frontier-scale open-source model with a 256k context window.

Kimi K2.6

model-card · v1.0.0

Frontier-scale open-source 1T parameter model with a 262.1k context window, designed for agentic workloads.

MeloTTS

model-card · v1.0.0

High-quality multi-lingual text-to-speech library.

Hermes 2 Pro Mistral 7B

model-card · v1.0.0

Upgraded, retrained version of Nous Hermes 2 with function calling and JSON mode capabilities.

Nemotron 3 120B A12B

model-card · v1.0.0

Hybrid MoE model with leading accuracy for multi-agent applications.

GPT OSS 120B

model-card · v1.0.0

Open-weight model designed for production, general-purpose, high-reasoning use cases.

GPT OSS 20B

model-card · v1.0.0

Open-weight model optimized for lower latency and local or specialized use cases.

Whisper

model-card · v1.0.0

General-purpose speech recognition model supporting multilingual recognition, speech translation, and language identification.

Whisper Large v3 Turbo

model-card · v1.0.0

Pre-trained model for automatic speech recognition (ASR) and speech translation.

Whisper Tiny EN

model-card · v1.0.0

English-only version of the Whisper Tiny model trained on speech recognition.

PLaMo Embedding 1B

model-card · v1.0.0

Japanese text embedding model that converts Japanese text input into numerical vectors for information retrieval, text classification, and clustering.

SmartTurn v2

model-card · v1.0.0

Open source community-driven native audio turn detection model in its second version.

Qwen2.5 Coder 32B Instruct

model-card · v1.0.0

Code-specific large language model supporting six mainstream model sizes from 0.5B to 32B parameters.

Qwen3 30B A3B FP8

model-card · v1.0.0

Latest generation large language model with groundbreaking advancements in reasoning, instruction-following, and agent capabilities.

Qwen3 Embedding 0.6B

model-card · v1.0.0

Latest Qwen family model specifically designed for text embedding and ranking.

QwQ 32B

model-card · v1.0.0

Medium-sized reasoning model capable of thinking deeply for hard problems; competitive with DeepSeek-R1 and o1-mini.

Stable Diffusion XL Base 1.0

model-card · v1.0.0

Diffusion-based text-to-image model that generates and modifies images based on text prompts.

UForm Gen2 Qwen 500M

model-card · v1.0.0

Small generative vision-language model primarily designed for image captioning and visual question answering.

GLM-4.7 Flash

model-card · v1.0.0

Fast and efficient multilingual text generation model with a 131,072 token context window supporting 100+ languages.