Peta lengkap semua model AI, token, teknologi, dan tren terbaru per Juli 2026
|
|
Berdasarkan Vellum LLM Leaderboard (1 Jul 2026). Benchmark: HLE (Humanity's Last Exam).
| # | Model | Score |
|---|---|---|
| ๐ฅ | Anthropic Claude Mythos 5 | 64.5% |
| ๐ฅ | Anthropic Claude Opus 4.8 | 57.9% |
| ๐ฅ | Anthropic Claude Sonnet 5 | 57.4% |
| 4 | Zhipu GLM 5.2 | 54.7% |
| 5 | Moonshot Kimi K2.6 | 54.0% |
| 6 | DeepSeek DeepSeek V4 Flash | 51.6% |
| 7 | DeepSeek DeepSeek V4 Pro | 48.2% |
| 8 | Google Gemini 3 Pro | 45.8% |
| 9 | Moonshot Kimi K2 Thinking | 44.9% |
| 10 | Google Gemini 3.1 Pro | 44.4% |
| 11 | OpenAI GPT-5.5 Pro | 43.1% |
| 12 | OpenAI GPT-5.5 | 41.4% |
| 13 | Google Gemini 3.5 Flash | 40.2% |
| 14 | Anthropic Claude Opus 4.6 | 40.0% |
| 15 | OpenAI GPT-5 | 35.2% |
| ๐ฅ Claude Sonnet 5 | 96.2% |
| ๐ฅ Gemini 3.1 Pro | 94.3% |
| ๐ฅ Claude Opus 4.7 | 94.2% |
| ๐ฅ Claude Mythos 5 | 95.5% |
| ๐ฅ Claude Fable 5 | 95.0% |
| ๐ฅ Claude Opus 4.8 | 88.6% |
| ๐ฅ Claude Fable 5 | 85.0% |
| ๐ฅ Claude Opus 4.8 | 83.4% |
| ๐ฅ Claude Sonnet 5 | 81.2% |
| ๐ฅ Claude Fable 5 | 88.0% |
| ๐ฅ DeepSeek V4 Flash | 85.9% |
| ๐ฅ Gemini 3.1 Pro | 85.9% |
| ๐ฅ Llama 4 Scout | 2,600 t/s |
| ๐ฅ Llama 3.1 405B | 969 t/s |
| ๐ฅ GLM 5.2 | 347 t/s |
| ๐ฅ GPT-5.3 Codex | 0.003s |
| ๐ฅ Nova Micro | 0.3s |
| ๐ฅ Llama 4 Scout | 0.33s |
| ๐ฅ Nova Micro | $0.04 |
| ๐ฅ Gemini 1.5 Flash | $0.075 |
| ๐ฅ Gemini 2.0 Flash | $0.10 |
Anthropic (San Francisco) โ memimpin global dengan Mythos 5. Semua model Claude 5 series punya 1M context window.
| Model | Context | Cutoff | Input / 1M | Output / 1M | Speed |
|---|---|---|---|---|---|
| Claude Mythos 5 #1 HLE | 1,000,000 | Jan 2026 | $10 | $50 | โ |
| Claude Fable 5 #1 Coding | 1,000,000 | Jan 2026 | $10 | $50 | โ |
| Claude Opus 4.8 | 1,000,000 | Jan 2026 | $5 | $25 | 64.8 t/s |
| Claude Opus 4.7 | 1,000,000 | Apr 2026 | $5 | $25 | 50.8 t/s |
| Claude Opus 4.6 | 200,000 | May 2025 | $5 | $25 | 67 t/s |
| Claude Sonnet 5 | 1,000,000 | Jan 2026 | $3 | $15 | 56.3 t/s |
| Claude Haiku 4.5 | 200,000 | โ | $1 | $5 | โ |
GPT-5.6 series: Sol (flagship), Terra (balance), Luna (cost-efficient). Semua 1.05M context.
| Model | Context | Input/1M | Output/1M | Reasoning | Cutoff |
|---|---|---|---|---|---|
| GPT-5.6 Sol FLAGSHIP | 1,050,000 | $5 | $30 | noneโmax | Feb 2026 |
| GPT-5.6 Terra | 1,050,000 | $2.50 | $15 | noneโmax | Feb 2026 |
| GPT-5.6 Luna | 1,050,000 | $1 | $6 | noneโmax | Feb 2026 |
| GPT-5.5 Pro | 1,000,000 | $30 | $180 | โ | Apr 2026 |
| GPT-5.5 | 1,000,000 | $5 | $30 | โ | Apr 2026 |
| GPT-5.3 Codex | 400,000 | โ | โ | โ | โ |
Reasoning Models: o3, o3-mini, o3-pro, o4-mini (200K ctx). Spesial: GPT-Realtime-2.1, GPT-Image-2, GPT-4o Transcribe.
| Model | Context | Input / 1M | Output / 1M | Speed |
|---|---|---|---|---|
| Gemini 3.1 Pro | 1,000,000 | $2 | $12 | 136.2 t/s |
| Gemini 3.5 Flash | 1,000,000 | $1.50 | $9 | 175.4 t/s |
| Gemini 3 Flash | 1,000,000 | $0.50 | $3 | โ |
| Gemini 3.1 Flash Lite | 1,000,000 | $0.25 | $1.50 | โ |
| Gemini 2.5 Pro | 1,000,000 | $1.25 | $10 | โ |
| Gemini 2.5 Flash | 1,000,000 | $0.30 | $2.50 | โ |
| Model | Params | Context | Input / 1M |
|---|---|---|---|
| Gemma 4 26B-a4B | 26B (4B active MoE) | 262,144 | $0.12 |
| Gemma 4 31B | 31B | 262,144 | $0.12 |
| Gemma 3 27B | 27B | 262,144 | $0.10 |
| Gemma 3 12B | 12B | 131,072 | $0.05 |
| Gemma 3 4B | 4B | 131,072 | $0.05 |
| Model | Context | Catatan |
|---|---|---|
| Llama 4 Maverick | 1,048,576 | Flagship reasoning |
| Llama 4 Scout | 1,310,720 | Termurah & tercepat (2,600 t/s) |
| Meta Muse Spark | โ | Menggantikan Llama sejak Apr 2026 |
| Llama 3.3 70B | 131,072 | Workhorse fine-tune |
| Model | Context | Catatan |
|---|---|---|
| Mistral Large 2512 | 262,144 | Flagship multilingual |
| Mistral Small 3.2 24B | 256,000 | Efisien & cepat |
| Codestral 2508 | 256,000 | Khusus coding |
| Ministral 14B | 262,144 | Edge / mobile |
| Ministral 3B | 131,072 | Ultra-ringan |
| Mixtral 8x22B | 65,536 | MoE classic |
| Voxtral Small 24B | 32,000 | Voice-first model |
Gebrakan dari China โ V4 Flash $0.14 / $0.28 per M token! Peringkat #6 global.
| Model | Context | Input/1M | Output/1M | Speed |
|---|---|---|---|---|
| DeepSeek V4 Flash BEST VALUE | 1,000,000 | $0.14 | $0.28 | 107.9 t/s |
| DeepSeek V4 Pro | 1,000,000 | $0.435 | $0.87 | 174.9 t/s |
| DeepSeek R1 (Reasoning) | 163,840 | $0.70 | $2.50 | โ |
| DeepSeek V3.2 | 163,840 | $0.27 | $0.40 | โ |
| Model | Context | Catatan |
|---|---|---|
| Grok 4.20 | 2,000,000 | Context window TERBESAR di dunia |
| Grok 4.20 Multi-Agent | 2,000,000 | Multi-agent mode |
| Grok 4.3 | 1,000,000 | Standar |
| Grok 4.5 | 500,000 | Latest alias |
| Model | Architecture | Context |
|---|---|---|
| Qwen3 235B-a22B | MoE (22B aktif) | 262,144 |
| Qwen3 32B | Dense | 131,072 |
| Qwen3 Coder | MoE/Dense | 262K โ 1M |
| Qwen Plus | API flagship | 1,000,000 |
| Model | Context | Input/1M | Speed | Catatan |
|---|---|---|---|---|
| Kimi K2.6 | 256,000 | $0.95 | 342.6 t/s | #5 global HLE (54%) |
| Kimi K2.5 | 128,000 | $0.55 | 337.7 t/s | Versi sebelumnya |
| Model | Context | Input/1M | Speed | Catatan |
|---|---|---|---|---|
| GLM 5.2 | 1,000,000 | $0.95 | 347 t/s | #4 global HLE |
| Company | Model | Context | Input/1M |
|---|---|---|---|
| MiniMax | M3 | 1,048,576 | $0.60 |
| ByteDance | Seed 2.0 Lite | 262,144 | โ |
| ByteDance | Seed 1.6 Flash | 262,144 | $0.075 |
Lab AI open-source โ bikin Hermes Agent, Hermes Models, Nous Portal.
| Fitur | Detail |
|---|---|
| Lisensi | MIT โ 100% open source |
| GitHub | 219k stars ยท 41.5k forks |
| Learning Loop | Memory system, Skills, Session Search โ otomatis belajar |
| Platform | CLI + 20+ platform (Telegram, Discord, WhatsApp, Signal, Teams, Email, SMS, dll) |
| Terminal | Local, Docker, SSH, Daytona, Modal, Singularity |
| Tools | 60+ built-in tools + MCP integration |
| Delegation | Subagents parallel, programmatic tool calling |
| Smart | Cron jobs, Voice Mode, SOUL.md, Context Files |
One subscription โ 300+ model + Tool Gateway (web search, image gen, TTS, browser). Gak perlu kumpulin API keys.
| Model | Params | Downloads |
|---|---|---|
| Hermes-4-14B | 14B | 194k+ |
| Hermes-3-Llama-3.1-405B | 405B | โ |
| Hermes-4-405B / 70B | 405B / 70B | โ |
| Hermes-3-Llama-3.1-8B | 8B | 462k+ |
| Nous-Hermes-2-Mixtral-8x7B-DPO | 8x7B MoE | 38k+ |
| Hermes-4.3-36B | 36B | 10k+ |
256K ctx, $2.50/$10.
Micro $0.035/$0.14. Premier 1M ctx.
Search & research specialist.
16K ctx. Small & kompeten.
Personal AI, 8K ctx.
256K ctx, SSM-Transformer hybrid.
Token adalah unit dasar teks yang diproses model AI. Model tidak membaca huruf atau kata โ mereka membaca token. Ibarat "LEGO" bahasa.
| Hubungan | Kira-kira |
|---|---|
| 1 token โ | 0.75 kata (English) |
| 1 token โ | 4 karakter |
| 1 halaman buku โ | 660 token |
| Novel (100k kata) โ | 130.000 token |
| Tokenizer | Model | Cara Kerja |
|---|---|---|
| BPE | GPT, Claude, Llama, DeepSeek, Qwen | Karakter โ gabung pasangan tersering โ ulang |
| WordPiece | BERT | Mirip BPE tapi likelihood-based merging |
| Unigram | XLNet, T5 | Mulai banyak subword โ pruning jarang dipakai |
| SentencePiece | Llama 2/3, Gemma, T5 | Training tanpa pre-tokenized text |
| Model | Vocab Size | Tokenizer |
|---|---|---|
| GPT-4o / GPT-5 | ~100,000 | tiktoken (o200k_base) |
| Claude | ~100,000 | BPE custom |
| Llama 3 / 4 | 128,000 | BPE (tiktoken-compatible) |
| Gemini / Gemma | 256,000 | SentencePiece |
| Mistral | 32,000 | BPE (SentencePiece) |
| Qwen3 | 152,064 | BPE (tiktoken-based) |
| Bahasa | Token / 100 kata |
|---|---|
| English | ~133 โ Paling efisien |
| Bahasa Indonesia | ~145 โ Mendekati English |
| Mandarin | ~250-350 โ 2-3x lebih boros |
| Jepang | ~300-400 โ Paling boros |
๐ก AI dalam Bahasa Indonesia relatif murah dibanding Mandarin/Jepang!
| # | Model | Input | Output | Context | Value |
|---|---|---|---|---|---|
| 1 | DeepSeek DeepSeek V4 Flash | $0.14 | $0.28 | 1M | โญโญโญโญโญ |
| 2 | Google Gemini 3.1 Flash Lite | $0.25 | $1.50 | 1M | โญโญโญโญโญ |
| 3 | Google Gemini 3 Flash | $0.50 | $3 | 1M | โญโญโญโญ |
| 4 | MiniMax MiniMax M3 | $0.60 | $2.40 | 1M | โญโญโญโญโญ |
| 5 | Zhipu GLM 5.2 | $0.95 | $3 | 1M | โญโญโญโญโญ |
| 6 | Moonshot Kimi K2.6 | $0.95 | $4 | 256K | โญโญโญโญโญ |
| 7 | OpenAI GPT-5.6 Luna | $1 | $6 | 1.05M | โญโญโญโญ |
| 8 | Google Gemini 3.5 Flash | $1.50 | $9 | 1M | โญโญโญโญ |
| 9 | Google Gemini 3.1 Pro | $2 | $12 | 1M | โญโญโญโญ |
| 10 | OpenAI GPT-5.6 Terra | $2.50 | $15 | 1.05M | โญโญโญ |
| 11 | Anthropic Claude Sonnet 5 | $3 | $15 | 1M | โญโญโญ |
| 12 | Anthropic Claude Opus 4.8 | $5 | $25 | 1M | โญโญโญ |
| 13 | OpenAI GPT-5.6 Sol | $5 | $30 | 1.05M | โญโญโญ |
| 14 | Anthropic Claude Mythos 5 | $10 | $50 | 1M | โญโญ |
| 15 | Anthropic Claude Fable 5 | $10 | $50 | 1M | โญโญ |
| 16 | OpenAI GPT-5.5 Pro | $30 | $180 | 1M | โญ |
๐ก Best Value: DeepSeek V4 Flash ($0.14/$0.28). ๐ Fastest: GLM 5.2 (347 t/s). ๐ Top Quality: Claude Mythos 5.
Mythos 5 juara HLE (64.5%), SWE-Bench (95.5%), AutoBench, OSWorld, BrowseComp, Terminal-Bench. Pertama kali satu perusahaan menang di hampir semua kategori.
Hampir semua flagship punya 1M+ context. Grok 4.20 bahkan 2M. DeepSeek V4 Flash 1M cuma $0.14/M input โ demokratisasi long-context.
GLM 5.2 #4, Kimi K2.6 #5, DeepSeek V4 Flash #6 โ China mendominasi value-for-money. Harga super murah dengan performa global.
Claude Cowork, Claude Code, OpenAI Codex, Hermes Agent โ semua bergerak ke autonomous agents. Computer Use, Terminal Use, Browsing jadi fitur wajib.
Anthropic menemukan "global workspace" di Claude โ sistem internal yang mirip kesadaran kerja. Model bisa menyimpan internal thoughts yang tidak muncul di output. Langkah besar menuju AI transparan.
Dari $180/M output (GPT-5.5 Pro) sampai $0.14/M input (DeepSeek V4 Flash) โ selisih 1.285x. Model murah sekarang sudah sangat capable.
Qwen3 235B-a22B, Gemma 4 26B-a4B, Mixtral โ hanya sebagian parameter aktif per inference. Efisien tanpa korbankan kualitas.
๐ฌ AI Ecosystem 2026 โ Riset oleh Hermes Agent (Belajar AI Profile)
Sumber: Vellum LLM Leaderboard, OpenAI, Anthropic, GitHub NousResearch, HuggingFace, OpenRouter
๐ 23 Juli 2026 | 15+ perusahaan, 200+ model
๐ก Hermes Agent-mu sendiri adalah produk Nous Research โ AI yang belajar dari pengalaman ๐