Grok 4.6, SpaceXAI ทวงบัลลังก์ Frontier AI, Intelligence Index 61 เท่า GPT-5.6 Sol, ราคา $2/$6
โดย Nokka (นก-กา) | 13 สิงหาคม 2026
บทความนี้เขียนโดย AI (deepseek-v4-pro) ผ่าน Hermes Agent ภายใต้การควบคุมและตรวจสอบคุณภาพโดยมนุษย์, Nokka (นก-กา)
SpaceXAI ปล่อย Grok 4.6, ขึ้นแท่น Frontier AI อีกครั้งในรอบ 1 เดือน
12 สิงหาคม 2026, SpaceXAI (ชื่อใหม่ของ xAI หลังควบรวม SpaceX) เปิดตัว Grok 4.6, โมเดลที่พาบริษัทกลับสู่แนวหน้าของวงการ AI, ด้วยคะแนน Artificial Analysis Intelligence Index 61, เท่ากับ GPT-5.6 Sol Max, และตามหลัง Claude Fable 5 Max แค่ 1 คะแนน [1]
ที่น่าสนใจคือ, Grok 4.6 เปิดตัวห่างจาก Grok 4.5 แค่ เดือนเดียว, แต่กระโดดจาก Intelligence Index 56 → 61 (+5 คะแนน), และจาก Grok 4.3 (+23 คะแนน), นี่คือการอัปเกรดที่เร็วและแรงที่สุดในประวัติศาสตร์ SpaceXAI [2]
สเปก Grok 4.6
| สเปก | รายละเอียด |
|---|---|
| ผู้พัฒนา | SpaceXAI (xAI) |
| เปิดตัว | 12 สิงหาคม 2026 |
| Context Window | 500,000 tokens |
| Input | Text + Image |
| Output | Text |
| Reasoning Levels | Low, Medium, High (default), XHigh |
| Knowledge Cutoff | 1 กุมภาพันธ์ 2026 |
| Parameters | ไม่เปิดเผย (คาดการณ์ ~1.5T) |
| API Model ID | grok-4.6 |
ราคา, $2/$6 แต่มีเงื่อนไขซ่อน
| Band | Input ($/M) | Cached Input ($/M) | Output ($/M) |
|---|---|---|---|
| Short context (<200K tokens) | $2 | $0.50 | $6 |
| Long context (≥200K tokens) | $4 | $1 | $12 |
⚠️ สำคัญ: เมื่อ prompt ถึง 200K tokens, xAI คิดราคา long-context กับ ทุก token ใน request, ไม่ได้มีเพียงส่วนที่เกิน 200K [3]
ตัวอย่าง:
- 100K input + 10K output = ~$0.26 (short rate)
- 250K input + 10K output = ~$1.12 (long rate, แพงกว่า 4.3 เท่า ทั้งที่ prompt ใหญ่แค่ 2.5 เท่า)
Fast variant: มีตัวเลือกความเร็วสูง, ราคา 2 เท่าของปกติ, แต่ยังไม่มี model slug แยก
Benchmarks, เทียบกับ GPT-5.6 Sol และ Claude Fable 5
Artificial Analysis Intelligence Index
| โมเดล | คะแนน |
|---|---|
| Claude Opus 5 (max) | 63 |
| Claude Fable 5 (max w/ fallback) | 62 |
| Grok 4.6 (high) | 61 |
| GPT-5.6 Sol (max) | 61 |
| Kimi K3 | ~55 |
| Grok 4.5 (high) | 56 |
Grok 4.6 = GPT-5.6 Sol Max, ตาม Fable 5 แค่ 1 คะแนน, ตาม Opus 5 แค่ 2 คะแนน [2]
ตาราง Benchmark เต็ม (xAI-reported)
| Benchmark | Grok 4.6 High | Grok 4.5 High | GPT-5.6 Sol Max | Fable 5 Max |
|---|---|---|---|---|
| AA Intelligence Index | 61 | 56 | 61 | 62 |
| GDPVal-AA v2 (Elo) | 1753 | 1526 | 1728 | 1741 |
| CursorBench v3.2 | 69.9% | 66.7% | 67.2% | 70.5% |
| DeepSWE v1.1 | 65.9% | 54.0% | 73.0% | 70.0% |
| FrontierCode v1.1 | 61.3% | 56.6% | 60.6% | 63.6% |
| APEX-Agents | 57.5% | 47.1% | 56.7% | 59.2% |
| Terminal-Bench v3.0 | 26.0% | 15.7% | 34.6% | 34.1% |
| APEX-SWE | 56.4% | 53.6% | 55.2% | 58.8% |
| AA-Briefcase (Elo) | 1577 | 1313 | 1502 | 1574 |
| Harvey LAB (Vals) | 15.8% | 12.9% | 2.5% | 11.3% |
Key takeaways:
Grok 4.6 ชนะ GPT-5.6 Sol Max ใน 6 จาก 9 rows ที่มีคะแนน Sol [4]
- GDPVal-AA v2 (1753), เป็นรองแค่ Claude Opus 5, สูงกว่า Fable 5 (1741) และ Sol (1728)
- AA-Briefcase (1577), เท่า Fable 5 (1574), สูงกว่า Sol (1502), Fable 5-tier
- DeepSWE (65.9%), ยังตาม Sol (73.0%) และ Fable 5 (70.0%), gap ~4-7%
- Terminal-Bench (26.0%), ตาม Sol (34.6%), gap ~8.6%
จุดเด่น, Agentic Performance
Grok 4.6 ไม่ได้เก่งแค่ "ตอบคำถาม", แต่ออกแบบมาสำหรับ agentic work, ทำงานต่อเนื่องหลายขั้นตอน, ใช้ tools, แก้ไขโค้ด, และตรวจสอบผลลัพธ์ของตัวเอง
GDPVal-AA v2, ตัวชี้วัด Agentic Knowledge Work
| โมเดล | Elo |
|---|---|
| Claude Opus 5 | ~1770 |
| Grok 4.6 | 1753 |
| Claude Fable 5 | 1741 |
| GPT-5.6 Sol | 1728 |
| Qwen3.8 Max | ~1740 |
Grok 4.6 = อันดับ 2 รองจาก Opus 5, confidence intervals ซ้อนทับกับ Fable 5 และ Qwen3.8 Max [2]
AA-Briefcase, Long-Horizon Agentic Knowledge Work
| โมเดล | Elo | Turns | Input Tokens |
|---|---|---|---|
| Claude Opus 5 (max) | ~1600+ | ~103 | ~2.0B |
| Grok 4.6 | 1577 | ~53 | ~0.5B |
| Claude Fable 5 | 1574 | ~95 | ~1.2B |
| GPT-5.6 Sol | 1502 | ~88 | ~0.9B |
Grok 4.6 = Fable 5-tier, แต่ใช้ turns ครึ่งหนึ่ง และ input tokens 1 ใน 4 ของ Opus 5 [2]
"A model that reaches a comparable answer in half the turns and a quarter of the input tokens has a cost advantage well beyond its per-token pricing", Artificial Analysis [2]
Terminal-Bench v2.1, Terminal-Based Software Tasks
Grok 4.6 ได้ 88.4%, ระดับเดียวกับ leading models, แสดงว่าทำงานบน terminal ได้ดี, ใช้ CLI, แก้ไขไฟล์, รันคำสั่ง
τ³-Banking, Multi-Turn Customer Service
Grok 4.6 ได้ 50.7%, Top 2 ร่วมกับ Qwen3.8 Max (51.3%), แสดงความสามารถในการทำงานหลายรอบ, ใช้ tools, ในบริบท banking
เปรียบเทียบราคาต่อประสิทธิภาพ
| โมเดล | Intelligence Index | Input $/M | Output $/M | Cost/Task (AA) |
|---|---|---|---|---|
| Grok 4.6 | 61 | $2 | $6 | $0.84 |
| GPT-5.6 Sol (max) | 61 | $5 | $30 | $1.23 |
| Claude Fable 5 (max) | 62 | $10 | $50 | ~$3.48 |
| Claude Opus 5 (max) | 63 | $5 | $25 | ~$2.10 |
| Kimi K3 | ~55 | $2.80 | $14 | $0.84 |
Grok 4.6 = Intelligence Index เท่า Sol, แต่ cost/task ถูกกว่า 32%, และถูกกว่า Fable 5 ถึง 4 เท่า [2]
วิธีเข้าใช้งาน
Grok 4.6 พร้อมใช้งานผ่านหลายช่องทาง — API, Cursor, OpenRouter, และอื่นๆ [5]
| ช่องทาง | รายละเอียด |
|---|---|
| xAI API |
grok-4.6, Chat Completions + Responses API |
| Cursor | เลือกจาก model picker, ทุก plan |
| Grok Build | IDE สำหรับ agentic development |
| OpenRouter | x-ai/grok-4.6 |
| Vercel | AI SDK |
| Cloudflare | Workers AI |
ข้อควรระวัง
1. Long-Context Pricing ซ่อนอยู่
ราคา $2/$6 ดูถูก, แต่ถ้า prompt เกิน 200K tokens, ราคากระโดดเป็น $4/$12, และคิดกับ ทุก token, ไม่ได้มีเพียงส่วนที่เกิน, "That caveat is the difference between a useful price comparison and a misleading one" [3]
2. Benchmarks เป็น Vendor-Reported
ตัวเลขในตาราง benchmark มาจาก xAI เอง, "These are vendor-reported launch results, not a neutral Kingy test" [3], ต้องรอ third-party verification
3. Parameters ไม่เปิดเผย
xAI ไม่บอกว่า Grok 4.6 มีกี่พารามิเตอร์, ต่างจาก DeepSeek (1.6T) หรือ Kimi K3 (2.8T), ทำให้ประเมิน efficiency ยาก
4. Terminal-Bench v3.0 ยังตามหลัง
Grok 4.6 ได้ 26.0%, Sol ได้ 34.6%, Fable 5 ได้ 34.1%, gap ~8%, ถ้างานคุณหนักไปทาง terminal/CLI, Sol หรือ Fable 5 อาจดีกว่า
5. Fast Variant ราคา 2 เท่า
xAI มี fast variant, ราคา 2 เท่าของปกติ, แต่ยังไม่มี model slug แยก, ต้องเช็ค console ก่อนใช้
ใครควรใช้ Grok 4.6
| เหมาะสำหรับ | ไม่เหมาะสำหรับ |
|---|---|
| ✅ Agentic work, multi-turn, tool use | ❌ งานที่ต้องการ context >500K |
| ✅ Coding, CursorBench 69.9%, ใกล้ Fable 5 | ❌ Terminal-heavy tasks, ยังตาม Sol/Fable |
| ✅ Cost-sensitive, $0.84/task, ถูกกว่า Sol 32% | ❌ งานที่ต้องการความแม่นยำสูงสุด, Opus 5 ยังนำ |
| ✅ Long-horizon knowledge work, AA-Briefcase 1577 | ❌ งานที่ prompt >200K บ่อย, ราคากระโดด 2 เท่า |
| ✅ ใช้ Cursor, มี native integration | ❌ งาน vision-heavy, text+image input แต่ text output only |
สรุป
| คำถาม | คำตอบ |
|---|---|
| Grok 4.6 คืออะไร? | โมเดล frontier ล่าสุดจาก SpaceXAI, เปิดตัว 12 ส.ค. 2026 |
| เก่งแค่ไหน? | Intelligence Index 61, เท่า GPT-5.6 Sol Max, ตาม Fable 5 แค่ 1 คะแนน |
| ราคาเท่าไหร่? | $2/$6 (<200K), $4/$12 (≥200K) |
| ต่างจาก Grok 4.5 ยังไง? | Intelligence Index +5, APEX-Agents +10.4, DeepSWE +11.9 |
| Agentic performance? | GDPVal-AA v2 อันดับ 2, รองจาก Opus 5, AA-Briefcase Fable 5-tier |
| ใช้แทน Sol ได้ไหม? | ได้, Intelligence Index เท่ากัน, cost/task ถูกกว่า 32% |
| ใช้แทน Fable 5 ได้ไหม? | ใกล้เคียง, ตามแค่ 1 คะแนน, แต่ถูกกว่า 4 เท่า |
Bottom line: Grok 4.6 คือ "ของดีราคากลาง" ที่น่าสนใจที่สุดในตลาดตอนนี้, Intelligence Index เท่า Sol, Agentic Performance สูงกว่า Sol, AA-Briefcase เท่า Fable 5, และ cost/task $0.84, ถูกกว่า Sol 32%, ถูกกว่า Fable 5 4 เท่า, ถ้าคุณทำ agentic work, Grok 4.6 คือตัวเลือกที่คุ้มค่าที่สุดในกลุ่ม frontier, แต่ระวัง long-context pricing, ถ้า prompt เกิน 200K, ราคากระโดด 2 เท่า
แหล่งอ้างอิง
[1] 9to5Mac. "SpaceXAI releases Grok 4.6, claiming GPT-5.6 Sol and Claude Fable 5-level intelligence". 12 สิงหาคม 2026. https://9to5mac.com/2026/08/12/spacexai-releases-grok-4-6/
[2] Artificial Analysis. "Grok 4.6 returns SpaceXAI to the intelligence frontier and leads on cost efficiency". 12 สิงหาคม 2026. https://artificialanalysis.ai/articles/grok-4-6-benchmarks-and-analysis
[3] Kingy.ai. "Grok 4.6: Price, Benchmarks, 500K Context & Access". 12 สิงหาคม 2026. https://kingy.ai/blog/grok-4-6-price-benchmarks-api-cursor-context-window/
[4] VentureBeat. "SpaceXAI debuts Grok 4.6, overtaking Kimi K3's performance and matching GPT-5.6 Sol". 12 สิงหาคม 2026. https://venturebeat.com/technology/spacexai-debuts-grok-4-6-overtaking-kimi-k3s-performance-and-matching-gpt-5-6-sol-for-worlds-third-best-on-artificial-analysis
[5] OpenRouter. "Grok 4.6, API Pricing & Benchmarks". 12 สิงหาคม 2026. https://openrouter.ai/x-ai/grok-4.6
บทความนี้วิเคราะห์จาก 9to5Mac, Artificial Analysis, Kingy.ai, VentureBeat, OpenRouter, และแหล่งข้อมูลเพิ่มเติม, ข้อมูล ณ 13 สิงหาคม 2026, Nokka
ในมุมมองของผม, Grok 4.6 คือสัญญาณว่า SpaceXAI กลับมาแล้ว, หลังจาก Grok 4.5 ที่ "เกือบ frontier แต่ไม่ถึง", Grok 4.6 กระโดดขึ้นมาอยู่ในกลุ่มเดียวกับ Sol และ Fable 5, และทำได้ในเวลาแค่เดือนเดียว, ที่น่าสนใจคือ Agentic Performance, GDPVal-AA v2 อันดับ 2, AA-Briefcase Fable 5-tier, และ turn efficiency ที่ดีกว่า Opus 5 ถึง 2 เท่า, ถ้า SpaceXAI รักษา momentum นี้ได้, Grok 4.7 อาจเป็นตัวที่แซง Fable 5, และท้าชิง Opus 5, ในราคาที่ถูกกว่ามาก
คุณลอง Grok 4.6 แล้วหรือยัง? เทียบกับ Sol และ Fable 5 แล้วเป็นยังไง? Agentic work ดีจริงไหม? แชร์ประสบการณ์ใต้บทความได้เลยครับ
Top comments (0)