Model Performance Leaderboard

Real-world latency, throughput, stability, and cost trends, providing objective references for enterprise-grade model selection.

Weekly Category Top

Large Language Model TOP 1
DeepSeek-V4-Pro
DeepSeek
Calls this week:14B
Growth Rate:+12%
Starting price:{{commonSite.currencySymbol}}0.06/M
Image Generation TOP 1
doubao-seedream-5.0
ByteDance
week's gen count:32K/imgs
Growth Rate:+12%
Multiplier::{{commonSite.currencySymbol}}0.04/imgs
Video Generation TOP 1
happyhorse-1.0
DeepSeek
week's gen time:12.5Ks
Growth Rate:+12%
Starting price:{{commonSite.currencySymbol}}0.145/s
TOP1
deepseek-v4-flash
DeepSeek
Weekly Usage:14B
Growth Rate:+15%
Input Price:{{commonSite.currencySymbol}}2.4/M
Output Price:{{commonSite.currencySymbol}}4.25/M
TOP2
deepseek-v3.2
DeepSeek
Weekly Usage:11B
Growth Rate:+12%
Input Price:{{commonSite.currencySymbol}}1.21/M
Output Price:{{commonSite.currencySymbol}}2.43/M
TOP3
deepseek-v4-pro
DeepSeek
Weekly Usage:10B
Growth Rate:+10%
Input Price:{{commonSite.currencySymbol}}1.63/M
Output Price:{{commonSite.currencySymbol}}3.65/M
Rank
Model
Provider
Type
Weekly Usage
Growth Rate
Latency(ms)
Input Price
Actions
deepseek-v4-flash
DeepSeek
Large Language Model
16.8B Tokens
+22%
78
{{commonSite.currencySymbol}} 0.13 / M
deepseek-v3.2
DeepSeek
Large Language Model
13.4B Tokens
+8%
102
{{commonSite.currencySymbol}} 0.215 / M
deepseek-v4-pro
DeepSeek
Large Language Model
10.6B Tokens
+17%
118
{{commonSite.currencySymbol}} 0.75 / M
4
doubao-seedream-5.0
ByteDance
Image generation
52K imgs
+26%
6,800
{{commonSite.currencySymbol}} 0.04 / imgs
5
kimi-k2.6
Moonshot AI
Large Language Model
7.8B Tokens
+34%
146
{{commonSite.currencySymbol}} 0.77 / M
6
glm-5.2
GLM AI
Large Language Model
6.9B Tokens
+38%
132
{{commonSite.currencySymbol}} 0.98 / M
7
Seedance-2.0-Fast
ByteDance
Video generation
24.8K s
+36%
31,000
{{commonSite.currencySymbol}} 5 / M
8
kimi-k3
Moonshot AI
Large Language Model
5.9B Tokens
+11%
188
{{commonSite.currencySymbol}} 3 / M

Model Performance Benchmarks (Past 7 Days)

Based on real-world production load testing data from the platform

Data Updated Every 6 Hours
{{ card.title }} {{ card.subTitle }}
{{ item.name }} {{ item.value }}
Supply & Demand Trends
Token Consumption (Last 7 Days)
Dynamic Pricing Trends (Last 7 Days)
Input Pricing $/1M

Scenario-Based Selection Guide

Coding & Development
DeepSeek-V4-Pro Claude 3.5 Sonnet GPT-40 Qwen-Max CodeLlama-70B

Low latency, high throughput, function calling, and long-context support for IDE plugins, automated coding, and code review.

View Details
Long-Context Summarization & Analysis
Claude 3.5 Sonnet Kimi-K2.6 Gemini 1.5 Pro GPT-4 Turbo Yi-34B-200K

Ultra-long context windows and high stability for legal documents, academic papers, and financial report analysis.

View Details
Multimodal / Vision Understanding
GPT-40 Gemini 1.5 Pro Claude 3.5 5onnet Qwen-VL-Max LLaVA-NeXT

Advanced vision understanding with multimodal input for automated labeling, content moderation, and chart analysis.

View Details
Semantic Search / Embeddings
text-embedding-ada-002 BAAI/bge-m3 Cohere Embed v3 Qwen-embedding Llama 3.2

High-precision embeddings for RAG, semantic retrieval, and recommendation systems.

View Details
Speech Recognition / Audio
Whisper Large v3 SenseVoice Parakeet-TDT Wav2Vec2 FunASR

Multilingual, high-accuracy real-time and batch processing for meeting transcription, subtitles, and voice assistants.

View Details
Reasoning / Logic / Mathematics
GPT-40 DeepSeek-V4-Pro Claude 3.5 Sonnet Qwen-Max Llama 3.1 405B

Advanced reasoning and mathematical capabilities for exam solving, data analysis, and intelligent decision-making.

View Details

Start Building with ApiSmart!

Switch with one line of code, millisecond response times. ApiSmart provides your foundational enterprise-grade AI infrastructure. Commencing in simplicity, culminating in infinity.

Get Started