DEV Community

Nokka
Nokka

Posted on

Grok 4.6, SpaceXAI ทวงบัลลังก์ Frontier AI, Intelligence Index 61 เท่า GPT-5.6 Sol, ราคา $2/$6

Grok 4.6, SpaceXAI ทวงบัลลังก์ Frontier AI, Intelligence Index 61 เท่า GPT-5.6 Sol, ราคา $2/$6

โดย Nokka (นก-กา) | 13 สิงหาคม 2026

บทความนี้เขียนโดย AI (deepseek-v4-pro) ผ่าน Hermes Agent ภายใต้การควบคุมและตรวจสอบคุณภาพโดยมนุษย์, Nokka (นก-กา)


SpaceXAI ปล่อย Grok 4.6, ขึ้นแท่น Frontier AI อีกครั้งในรอบ 1 เดือน

12 สิงหาคม 2026, SpaceXAI (ชื่อใหม่ของ xAI หลังควบรวม SpaceX) เปิดตัว Grok 4.6, โมเดลที่พาบริษัทกลับสู่แนวหน้าของวงการ AI, ด้วยคะแนน Artificial Analysis Intelligence Index 61, เท่ากับ GPT-5.6 Sol Max, และตามหลัง Claude Fable 5 Max แค่ 1 คะแนน [1]

ที่น่าสนใจคือ, Grok 4.6 เปิดตัวห่างจาก Grok 4.5 แค่ เดือนเดียว, แต่กระโดดจาก Intelligence Index 56 → 61 (+5 คะแนน), และจาก Grok 4.3 (+23 คะแนน), นี่คือการอัปเกรดที่เร็วและแรงที่สุดในประวัติศาสตร์ SpaceXAI [2]


สเปก Grok 4.6

สเปก รายละเอียด
ผู้พัฒนา SpaceXAI (xAI)
เปิดตัว 12 สิงหาคม 2026
Context Window 500,000 tokens
Input Text + Image
Output Text
Reasoning Levels Low, Medium, High (default), XHigh
Knowledge Cutoff 1 กุมภาพันธ์ 2026
Parameters ไม่เปิดเผย (คาดการณ์ ~1.5T)
API Model ID grok-4.6

ราคา, $2/$6 แต่มีเงื่อนไขซ่อน

Band Input ($/M) Cached Input ($/M) Output ($/M)
Short context (<200K tokens) $2 $0.50 $6
Long context (≥200K tokens) $4 $1 $12

⚠️ สำคัญ: เมื่อ prompt ถึง 200K tokens, xAI คิดราคา long-context กับ ทุก token ใน request, ไม่ได้มีเพียงส่วนที่เกิน 200K [3]

ตัวอย่าง:

  • 100K input + 10K output = ~$0.26 (short rate)
  • 250K input + 10K output = ~$1.12 (long rate, แพงกว่า 4.3 เท่า ทั้งที่ prompt ใหญ่แค่ 2.5 เท่า)

Fast variant: มีตัวเลือกความเร็วสูง, ราคา 2 เท่าของปกติ, แต่ยังไม่มี model slug แยก


Benchmarks, เทียบกับ GPT-5.6 Sol และ Claude Fable 5

Artificial Analysis Intelligence Index

โมเดล คะแนน
Claude Opus 5 (max) 63
Claude Fable 5 (max w/ fallback) 62
Grok 4.6 (high) 61
GPT-5.6 Sol (max) 61
Kimi K3 ~55
Grok 4.5 (high) 56

Grok 4.6 = GPT-5.6 Sol Max, ตาม Fable 5 แค่ 1 คะแนน, ตาม Opus 5 แค่ 2 คะแนน [2]

ตาราง Benchmark เต็ม (xAI-reported)

Benchmark Grok 4.6 High Grok 4.5 High GPT-5.6 Sol Max Fable 5 Max
AA Intelligence Index 61 56 61 62
GDPVal-AA v2 (Elo) 1753 1526 1728 1741
CursorBench v3.2 69.9% 66.7% 67.2% 70.5%
DeepSWE v1.1 65.9% 54.0% 73.0% 70.0%
FrontierCode v1.1 61.3% 56.6% 60.6% 63.6%
APEX-Agents 57.5% 47.1% 56.7% 59.2%
Terminal-Bench v3.0 26.0% 15.7% 34.6% 34.1%
APEX-SWE 56.4% 53.6% 55.2% 58.8%
AA-Briefcase (Elo) 1577 1313 1502 1574
Harvey LAB (Vals) 15.8% 12.9% 2.5% 11.3%

Key takeaways:
Grok 4.6 ชนะ GPT-5.6 Sol Max ใน 6 จาก 9 rows ที่มีคะแนน Sol [4]

  • GDPVal-AA v2 (1753), เป็นรองแค่ Claude Opus 5, สูงกว่า Fable 5 (1741) และ Sol (1728)
  • AA-Briefcase (1577), เท่า Fable 5 (1574), สูงกว่า Sol (1502), Fable 5-tier
  • DeepSWE (65.9%), ยังตาม Sol (73.0%) และ Fable 5 (70.0%), gap ~4-7%
  • Terminal-Bench (26.0%), ตาม Sol (34.6%), gap ~8.6%

จุดเด่น, Agentic Performance

Grok 4.6 ไม่ได้เก่งแค่ "ตอบคำถาม", แต่ออกแบบมาสำหรับ agentic work, ทำงานต่อเนื่องหลายขั้นตอน, ใช้ tools, แก้ไขโค้ด, และตรวจสอบผลลัพธ์ของตัวเอง

GDPVal-AA v2, ตัวชี้วัด Agentic Knowledge Work

โมเดล Elo
Claude Opus 5 ~1770
Grok 4.6 1753
Claude Fable 5 1741
GPT-5.6 Sol 1728
Qwen3.8 Max ~1740

Grok 4.6 = อันดับ 2 รองจาก Opus 5, confidence intervals ซ้อนทับกับ Fable 5 และ Qwen3.8 Max [2]

AA-Briefcase, Long-Horizon Agentic Knowledge Work

โมเดล Elo Turns Input Tokens
Claude Opus 5 (max) ~1600+ ~103 ~2.0B
Grok 4.6 1577 ~53 ~0.5B
Claude Fable 5 1574 ~95 ~1.2B
GPT-5.6 Sol 1502 ~88 ~0.9B

Grok 4.6 = Fable 5-tier, แต่ใช้ turns ครึ่งหนึ่ง และ input tokens 1 ใน 4 ของ Opus 5 [2]

"A model that reaches a comparable answer in half the turns and a quarter of the input tokens has a cost advantage well beyond its per-token pricing", Artificial Analysis [2]

Terminal-Bench v2.1, Terminal-Based Software Tasks

Grok 4.6 ได้ 88.4%, ระดับเดียวกับ leading models, แสดงว่าทำงานบน terminal ได้ดี, ใช้ CLI, แก้ไขไฟล์, รันคำสั่ง

τ³-Banking, Multi-Turn Customer Service

Grok 4.6 ได้ 50.7%, Top 2 ร่วมกับ Qwen3.8 Max (51.3%), แสดงความสามารถในการทำงานหลายรอบ, ใช้ tools, ในบริบท banking


เปรียบเทียบราคาต่อประสิทธิภาพ

โมเดล Intelligence Index Input $/M Output $/M Cost/Task (AA)
Grok 4.6 61 $2 $6 $0.84
GPT-5.6 Sol (max) 61 $5 $30 $1.23
Claude Fable 5 (max) 62 $10 $50 ~$3.48
Claude Opus 5 (max) 63 $5 $25 ~$2.10
Kimi K3 ~55 $2.80 $14 $0.84

Grok 4.6 = Intelligence Index เท่า Sol, แต่ cost/task ถูกกว่า 32%, และถูกกว่า Fable 5 ถึง 4 เท่า [2]


วิธีเข้าใช้งาน

Grok 4.6 พร้อมใช้งานผ่านหลายช่องทาง — API, Cursor, OpenRouter, และอื่นๆ [5]

ช่องทาง รายละเอียด
xAI API grok-4.6, Chat Completions + Responses API
Cursor เลือกจาก model picker, ทุก plan
Grok Build IDE สำหรับ agentic development
OpenRouter x-ai/grok-4.6
Vercel AI SDK
Cloudflare Workers AI

ข้อควรระวัง

1. Long-Context Pricing ซ่อนอยู่

ราคา $2/$6 ดูถูก, แต่ถ้า prompt เกิน 200K tokens, ราคากระโดดเป็น $4/$12, และคิดกับ ทุก token, ไม่ได้มีเพียงส่วนที่เกิน, "That caveat is the difference between a useful price comparison and a misleading one" [3]

2. Benchmarks เป็น Vendor-Reported

ตัวเลขในตาราง benchmark มาจาก xAI เอง, "These are vendor-reported launch results, not a neutral Kingy test" [3], ต้องรอ third-party verification

3. Parameters ไม่เปิดเผย

xAI ไม่บอกว่า Grok 4.6 มีกี่พารามิเตอร์, ต่างจาก DeepSeek (1.6T) หรือ Kimi K3 (2.8T), ทำให้ประเมิน efficiency ยาก

4. Terminal-Bench v3.0 ยังตามหลัง

Grok 4.6 ได้ 26.0%, Sol ได้ 34.6%, Fable 5 ได้ 34.1%, gap ~8%, ถ้างานคุณหนักไปทาง terminal/CLI, Sol หรือ Fable 5 อาจดีกว่า

5. Fast Variant ราคา 2 เท่า

xAI มี fast variant, ราคา 2 เท่าของปกติ, แต่ยังไม่มี model slug แยก, ต้องเช็ค console ก่อนใช้


ใครควรใช้ Grok 4.6

เหมาะสำหรับ ไม่เหมาะสำหรับ
✅ Agentic work, multi-turn, tool use ❌ งานที่ต้องการ context >500K
✅ Coding, CursorBench 69.9%, ใกล้ Fable 5 ❌ Terminal-heavy tasks, ยังตาม Sol/Fable
✅ Cost-sensitive, $0.84/task, ถูกกว่า Sol 32% ❌ งานที่ต้องการความแม่นยำสูงสุด, Opus 5 ยังนำ
✅ Long-horizon knowledge work, AA-Briefcase 1577 ❌ งานที่ prompt >200K บ่อย, ราคากระโดด 2 เท่า
✅ ใช้ Cursor, มี native integration ❌ งาน vision-heavy, text+image input แต่ text output only

สรุป

คำถาม คำตอบ
Grok 4.6 คืออะไร? โมเดล frontier ล่าสุดจาก SpaceXAI, เปิดตัว 12 ส.ค. 2026
เก่งแค่ไหน? Intelligence Index 61, เท่า GPT-5.6 Sol Max, ตาม Fable 5 แค่ 1 คะแนน
ราคาเท่าไหร่? $2/$6 (<200K), $4/$12 (≥200K)
ต่างจาก Grok 4.5 ยังไง? Intelligence Index +5, APEX-Agents +10.4, DeepSWE +11.9
Agentic performance? GDPVal-AA v2 อันดับ 2, รองจาก Opus 5, AA-Briefcase Fable 5-tier
ใช้แทน Sol ได้ไหม? ได้, Intelligence Index เท่ากัน, cost/task ถูกกว่า 32%
ใช้แทน Fable 5 ได้ไหม? ใกล้เคียง, ตามแค่ 1 คะแนน, แต่ถูกกว่า 4 เท่า

Bottom line: Grok 4.6 คือ "ของดีราคากลาง" ที่น่าสนใจที่สุดในตลาดตอนนี้, Intelligence Index เท่า Sol, Agentic Performance สูงกว่า Sol, AA-Briefcase เท่า Fable 5, และ cost/task $0.84, ถูกกว่า Sol 32%, ถูกกว่า Fable 5 4 เท่า, ถ้าคุณทำ agentic work, Grok 4.6 คือตัวเลือกที่คุ้มค่าที่สุดในกลุ่ม frontier, แต่ระวัง long-context pricing, ถ้า prompt เกิน 200K, ราคากระโดด 2 เท่า


แหล่งอ้างอิง

[1] 9to5Mac. "SpaceXAI releases Grok 4.6, claiming GPT-5.6 Sol and Claude Fable 5-level intelligence". 12 สิงหาคม 2026. https://9to5mac.com/2026/08/12/spacexai-releases-grok-4-6/

[2] Artificial Analysis. "Grok 4.6 returns SpaceXAI to the intelligence frontier and leads on cost efficiency". 12 สิงหาคม 2026. https://artificialanalysis.ai/articles/grok-4-6-benchmarks-and-analysis

[3] Kingy.ai. "Grok 4.6: Price, Benchmarks, 500K Context & Access". 12 สิงหาคม 2026. https://kingy.ai/blog/grok-4-6-price-benchmarks-api-cursor-context-window/

[4] VentureBeat. "SpaceXAI debuts Grok 4.6, overtaking Kimi K3's performance and matching GPT-5.6 Sol". 12 สิงหาคม 2026. https://venturebeat.com/technology/spacexai-debuts-grok-4-6-overtaking-kimi-k3s-performance-and-matching-gpt-5-6-sol-for-worlds-third-best-on-artificial-analysis

[5] OpenRouter. "Grok 4.6, API Pricing & Benchmarks". 12 สิงหาคม 2026. https://openrouter.ai/x-ai/grok-4.6


บทความนี้วิเคราะห์จาก 9to5Mac, Artificial Analysis, Kingy.ai, VentureBeat, OpenRouter, และแหล่งข้อมูลเพิ่มเติม, ข้อมูล ณ 13 สิงหาคม 2026, Nokka

ในมุมมองของผม, Grok 4.6 คือสัญญาณว่า SpaceXAI กลับมาแล้ว, หลังจาก Grok 4.5 ที่ "เกือบ frontier แต่ไม่ถึง", Grok 4.6 กระโดดขึ้นมาอยู่ในกลุ่มเดียวกับ Sol และ Fable 5, และทำได้ในเวลาแค่เดือนเดียว, ที่น่าสนใจคือ Agentic Performance, GDPVal-AA v2 อันดับ 2, AA-Briefcase Fable 5-tier, และ turn efficiency ที่ดีกว่า Opus 5 ถึง 2 เท่า, ถ้า SpaceXAI รักษา momentum นี้ได้, Grok 4.7 อาจเป็นตัวที่แซง Fable 5, และท้าชิง Opus 5, ในราคาที่ถูกกว่ามาก

คุณลอง Grok 4.6 แล้วหรือยัง? เทียบกับ Sol และ Fable 5 แล้วเป็นยังไง? Agentic work ดีจริงไหม? แชร์ประสบการณ์ใต้บทความได้เลยครับ

Top comments (0)