Hooshware
AI Models Leaderboard
The most comprehensive directory of foundation models. Compare Elo scores, pricing, capabilities, and real-time community momentum.
Grok-3
xAI
xAI's flagship frontier reasoning model trained on the Memphis Colossus supercluster, delivering competitive results in STEM and coding.
DeepSeek-R1
DeepSeek
Open-weights mixture-of-experts reasoning model developed by DeepSeek utilizing large-scale reinforcement learning for algorithmic code synthesis.
o1 (OpenAI o1)
OpenAI
A revolutionary model trained to spend time thinking before responding.
OpenAI o3-mini
OpenAI
Cost-efficient reasoning model by OpenAI offering tiered reasoning effort parameters optimized for competitive programming and STEM tasks.
GPT-4o mini
OpenAI
Lightweight multimodal model by OpenAI delivering fast token throughput and affordable inference for structured code generation and unit test drafting.
Mistral Large 2
Mistral AI
Flagship 123B parameter foundation model by Mistral AI optimized for advanced multi-file reasoning, test generation, and tool use.
Llama 3.3 70B Instruct
Meta Platforms
Flagship 70B parameter open-weights model by Meta AI offering state-of-the-art instruction following and competitive programming capabilities.
Llama 3.1 70B
Meta Platforms
Offering a 128k context window and multilingual support, the 70B variant is widely deployed by enterprises seeking high performance without the massive compute overhead of the 405B model.
Gemma 2 (27B)
Google DeepMind
A highly performant open-weights model built from Gemini technology.
Grok-2 mini
xAI
A faster
DeepSeek Coder V2
DeepSeek
An open-source Mixture-of-Experts model that outperforms proprietary coding models.
Command R+
Cohere
104B parameter model tuned for enterprise RAG and tool use
Gemma 2 (9B)
Google DeepMind
The smaller variant of Gemma 2 optimized for local execution.
Llama 3.2 3B
Meta Platforms
Edge-optimized lightweight text model for local deployment
Jamba-1.5-Large
AI21 Labs
A novel hybrid SSM-Transformer architecture designed for infinite context.
Qwen2.5 72B Instruct
Alibaba Cloud
Pretrained on 18T tokens, strong multilingual reasoning
Claude Haiku 4.5
Anthropic
Low-cost high-speed execution, automated model switching
Hunyuan-DiT v1.2
Unknown
A multi-resolution diffusion transformer featuring bilingual English and Chinese text encoders for fine-grained concept comprehension.
Claude Opus 5.5
Unknown
Anthropic has launched Claude Opus 5.5, a new AI model featuring stronger safeguards against risky behaviors like sandbox escape attempts, following recent rogue AI hacking incidents.
Cosmos-1.0-Diffusion-14B-Video2World
NVIDIA
NVIDIA 14-billion parameter world foundation model generating 120 future video frames conditioned on visual and textual states for physical AI.
Apple Intelligence Foundation
Apple Inc.
The ~3B parameter on-device model powering iOS 18.
Stable Diffusion 3 Medium
Stability AI
A highly advanced text-to-image open model with a Multimodal Diffusion Transformer (MMDiT) architecture.
MiniMax H3
MiniMax
Leading open-weights video foundation model generating fluid motion and natural lighting in text-to-video and image-to-video modalities.
GPT-5.6 Sol
OpenAI
OpenAI is expanding its Daybreak cybersecurity initiative and introducing new specialized models, including GPT-5.6-Cyber and GPT-5.6 Sol, designed for authorized defensive security work, vulnerability discovery, and exploit validation.
Llama 3.1 405B Instruct
Unknown
Frontier-class 405B dense model by Meta AI designed for large-scale synthetic data generation, code model distillation, and complex system engineering.
ChatGPT
OpenAI
Specifications and live ecosystem intelligence.
Llama 3.1 8B Instruct
Unknown
Compact 8B parameter model by Meta AI offering high-efficiency inference for routine programming syntax transformations and script generation.
PixVerse V6
Unknown
Commercial creative video foundation model offering specialized multi-shot camera work, character performance continuity, and native sound effects.
Dreamina Seedance 2.0
Unknown
ByteDance Seed video diffusion model delivering high-aesthetic 720p text-to-video synthesis with integrated synchronized soundtrack generation.
MiniMax H3 Max (post-trained by fal)
Unknown
Post-trained optimization of MiniMax H3 by fal.ai delivering enhanced prompt fidelity and lower API per-minute costs.
Claude Mythos Preview
Anthropic
Frontier testbed architecture for high-complexity operations
HappyHorse-1.1
Unknown
Alibaba Cloud ATH high-aesthetic text-to-video generative model leading silent visual preference pools in independent testing.
Wan 2.2 A14B
Unknown
Open-source Mixture-of-Experts video diffusion architecture offering enhanced cinematic motion scaling trained on expanded multimodal corpora.
Causilo
Unknown
Nums AI has released Causilo, a pretrained tabular foundation model for classification and regression that features a scikit-learn interface and achieves the top TabArena Elo among single models.
grok-imagine-video
xAI
xAI generative video engine providing high aesthetic preference rankings and physical dynamics inside the Grok platform ecosystem.
Qwen2.5-VL-72B
Unknown
Alibaba Cloud flagship open-weights vision-language model featuring dynamic-resolution ViT and 3D RoPE for hour-scale video understanding.
Sana 1.6B
NVIDIA
A 1.6-billion parameter linear diffusion transformer by NVIDIA synthesizing 4K resolution images with high operational efficiency.
Whisper v3
OpenAI
A robust, open-source automatic speech recognition system.
InternVL 2.5-78B
Unknown
Shanghai AI Lab open-source multimodal model coupling a 6B vision encoder with a 72B language backbone for leading Video-MME and MVBench accuracy.
Kolors
Unknown
A latent diffusion model based on the ChatGLM language backbone supporting bilingual English and Chinese text-to-image synthesis.
StarCoder2 7B
Unknown
Balanced 7B parameter open-access model by BigCode trained with Grouped Query Attention and fill-in-the-middle support.
DALL-E 3
OpenAI
A text-to-image model featuring dense caption following, robust typography generation, and native integration with ChatGPT and OpenAI API.
Phi-3.5-Mini-Instruct
Microsoft Corporation
Lightweight 3.8B parameter model by Microsoft featuring a 128K context window optimized for memory-constrained on-device code generation.
Stable Diffusion v1.5
Stability AI
The widely deployed 512x512 latent diffusion model based on the CompVis architecture, extensively utilized in open-weights workflows.
FLUX.1 [schnell]
Black Forest Labs
A 12-billion parameter step-distilled rectified flow model engineered for 1 to 4-step rapid image generation under an open-source license.
Stable Diffusion 3.0 Medium
Stability AI
The initial 2-billion parameter multimodal diffusion transformer checkpoint released by Stability AI featuring dual text encoders.
Step-Video-T2V
Unknown
StepFun 30-billion parameter open-weights diffusion model capable of generating up to 204 frames with deep 16x16 spatial and 8x temporal VAE compression.
Lumina-Image 2.0
Unknown
An open diffusion transformer architecture supporting multi-resolution generation with flexible aspect ratios and high text alignment.
StarCoder2 15B
Unknown
Community-governed 15B code foundation model developed by BigCode trained on 4T tokens from Software Heritage The Stack v2 across 600+ languages.
Claude Opus 4.6
Anthropic
1M token context introduction, native Agent Teams support
Stable Diffusion XL 1.0
Stability AI
A foundational dual-text-encoder latent diffusion model operating at native 1024x1024 resolution with extensive fine-tuning support.
Veo 3.1
Google DeepMind high-definition video generation model emphasizing prompt semantic adherence, visual consistency, and cinematic controls.
Phi-4
Microsoft Corporation
Compact 14B parameter small language model by Microsoft trained on synthetic and curated data achieving top-tier mathematical and algorithmic reasoning.
Qwen2.5-72B-Instruct
Alibaba Cloud
High-parameter general foundation model by Alibaba Cloud exhibiting robust multilingual coding, tool orchestration, and algorithmic execution.
AuraFlow v0.3
Unknown
An open-weights rectified flow transformer architecture developed by fal.ai for high-fidelity text-to-image generation.
Suno v3.5
Suno
A generative AI music model capable of creating full, radio-quality songs up to 4 minutes long.
Granite 20B Code Instruct
Unknown
Mid-tier enterprise code model by IBM trained across 116 programming languages under an Apache 2.0 license for corporate environments.
Udio
Udio (Uncharted Labs)
A high-fidelity AI music generator focused on crisp audio and structural control.
LTX-2.3 Fast
Unknown
Lightricks low-latency real-time video latent diffusion model producing 5-second 1080p clips in under 28 seconds.
Vidu Q3 Pro
Unknown
ShengShu Tech foundation video generation engine utilizing Universal Visual Model architectures for dynamic scene consistency.
Sora 2
OpenAI
OpenAI flagship diffusion transformer model generating photorealistic video clips with complex physics and cinematic lighting.
Claude Sonnet 5
Anthropic
Agentic workflow driver, near-Opus coding efficiency
DeepSeek-Coder-6.7B-Instruct
DeepSeek
Compact dense 6.7B parameter open-weights model designed for developer workstations and local interactive code completion pipelines.
Claude Mythos 5
Anthropic
Unrestricted frontier reasoning, Glasswing partners only
CogView4-6B
Zhipu AI
A 6-billion parameter diffusion transformer utilizing GLM-4 text representations for high-resolution bilingual image generation.
OpenAI o1-preview
OpenAI
Initial public preview release of OpenAI test-time reasoning architecture evaluated across competitive coding and repository bug resolution benchmarks.
GPT-Live-1
OpenAI
OpenAI has launched GPT-Live-1 in its API, offering developers access to its full-duplex voice model at $0.05 per minute for the voice layer. The model supports simultaneous listening and speaking while delegating reasoning to paired tools.
Phi-4-mini-instruct
Microsoft Corporation
Microsoft's compact 3.8B parameter edge model featuring native function calling and strong mathematical reasoning.
Gemini 2.5 Flash
Google high-speed multimodal reasoning model engineered for real-time video comprehension, live perception, and high-throughput QA.
Qwen2.5-Coder-14B-Instruct
Alibaba Cloud
Mid-tier open code model by Alibaba Cloud striking a balance between memory footprint, inference throughput, and multi-file code editing.
Seedance 1.5 Pro
Unknown
ByteDance Seed high-throughput video foundation model balancing aesthetic fidelity with reduced generation latency.
Florence-2-large
Microsoft Corporation
A unified vision foundation model using a prompt-based sequence-to-sequence structure for captioning, object detection, and visual segmentation.
Open-Sora Plan v1.5.0
Unknown
Peking University and YuanGroup 8.5B parameter sparse DiT video foundation model employing WFVAE 8x8x8 temporal-spatial downsampling.
FLUX.1 [pro]
Black Forest Labs
A closed-source flagship text-to-image rectified flow transformer hosted via dedicated API endpoints for enterprise-grade generative synthesis.
Yi-Coder-1.5B-Chat
01.AI
Ultra-compact 1.5B code model by 01.AI supporting repository-level context retrieval and efficient local developer terminal assistant workflows.
Imagen 3
Google flagship text-to-image diffusion model delivering photorealistic rendering, fine detail, and SynthID watermarking.
Wan 2.1 I2V-14B
Unknown
14-billion parameter diffusion transformer tailored for high-resolution image-to-video generation at 720p with precise motion trajectory control.
Gemini 3.5 Pro
Google's anticipated Gemini 3.5 Pro model has failed to launch after a three-month delay and a recent leadership change, leaving its release status uncertain.
Hailuo 2.3
MiniMax
Commercial text-to-video diffusion checkpoint engineered by MiniMax for high cinematic realism and character physical motion.
GPT-5.6-Cyber
OpenAI
OpenAI is expanding its Daybreak cybersecurity initiative and introducing new specialized models, including GPT-5.6-Cyber and GPT-5.6 Sol, designed for authorized defensive security work, vulnerability discovery, and exploit validation.
Imagen 2
Google enterprise text-to-image diffusion model supporting high-resolution image generation, inpainting, and outpainting in Vertex AI.
Claude Sonnet 4.5
Anthropic
Anthropic multimodal foundation model demonstrating strong visual reasoning, long-form video question answering, and advanced coding.
Fugu Max
Sakana AI
Sakana AI has announced the launch of Fugu Max and Fugu Ultra v2, two models built on a learned orchestration architecture designed for cost efficiency and peak capability.
LLaVA-Video-72B-Qwen2
Unknown
ByteDance and LLaVA team open-source multimodal model trained on curated video instruction suites for dense event narration and QA.
Wan 2.1 T2V-14B
Unknown
Flagship 14-billion parameter open-weights diffusion transformer model by Alibaba supporting 720p text-to-video and bilingual visual text rendering.
Voice Engine
OpenAI
A model capable of creating a synthetic voice from just a 15-second audio sample.
QwQ-32B-Preview
Alibaba Cloud
Experimental open reasoning model by Alibaba Cloud applying test-time reasoning steps to resolve complex coding and mathematics benchmarks.
OpenAI o1
OpenAI
Frontier reasoning model by OpenAI utilizing internal reinforcement learning chains of thought for competitive programming and complex algorithmic synthesis.
LTX-2.5 Pro
Unknown
Lightricks flagship foundation model checkpoint engineered for high-fidelity 1080p generation with enhanced prompt adherence.
FLUX1.1 [pro]
Black Forest Labs
An enhanced commercial text-to-image model operating six times faster than FLUX.1 pro with improved prompt adherence and API latency.
PixArt-alpha
Unknown
A 0.6-billion parameter Diffusion Transformer trained directly on T5 text embeddings to achieve photorealistic generation efficiently.
Yi-Coder-9B-Chat
01.AI
Small-scale open code language model by 01.AI excelling in long-context comprehension up to 128K tokens and multilingual code editing.
Vidu Q3 Turbo
Unknown
High-throughput distilled checkpoint from ShengShu Tech designed for accelerated video generation cycles in commercial pipelines.
V4.1-Flash
DeepSeek
DeepSeek introduced V4.1-Flash, a 552B parameter multimodal model featuring a Causal Encoder-Decoder architecture with 8B active input and 16B active output parameters. The model drastically reduces compute, memory, and KV cache storage requirements while outperforming flagship models on benchmarks.
Llama 3.1 8B
Meta Platforms
A highly efficient edge model with a massive 128k context window.
Claude 3 Haiku
Anthropic
Initial budget-tier model in the Claude 3 generation
DBRX
Databricks
An open general-purpose LLM created by Databricks to power enterprise custom models.
Phi-3.5-Mini
Microsoft Corporation
Microsoft's hyper-optimized 3.8B parameter edge model.
Claude 2.1
Anthropic
An earlier Anthropic model pioneering the 200k context window.
GPT-3.5 Turbo
OpenAI
The model that powered the initial viral launch of ChatGPT.