Grok Build 0.1 is xAI’s fast coding model trained specifically for agentic software engineering workflows. It supports text and image inputs with text output, and is optimized for interactive coding...
Tested on 8 benchmarks with 68.9% average. Top scores: LiveBench — Mathematics (78.4%), LiveBench — Reasoning (76.4%), LiveBench — Language (72.5%).
Llama 3.3 70B Instruct scores 75.9 (98% as good) at $0.10/1M input · 90% cheaper
Regularly refreshed coding problems that avoid data contamination. New problems added monthly to prevent memorization.
LiveBench coding tasks that require multi-step reasoning and tool use. Tests planning and execution of complex coding workflows.
Regularly refreshed reasoning problems testing logical deduction, spatial reasoning, and analytical thinking.
Fresh data analysis tasks testing ability to interpret tables, charts, and statistical data.
Regularly updated math problems that test numerical reasoning, algebra, calculus, and combinatorics.
- Typemultimodal
- Context256K tokens (~128 books)
- ReleasedMay 2026
- LicenseProprietary
- StatusActive
- Cost / Message~$0.004