xAI introduces Grok 4.6 with long-running agents and interactive AI capabilities


xAI has introduced Grok 4.6, its latest AI model focused on long-running agents and interactive and visual work. The model builds on Grok 4.5 and is designed to handle complex tasks across multiple steps, including researching topics, analyzing information, working across codebases, and turning ideas into applications or other work artifacts.

Grok 4.6 achieves frontier intelligence across several agentic coding and knowledge-work benchmarks. It matches GPT-5.6 Sol on the Artificial Analysis Intelligence Index, a composite score based on nine benchmarks.

Training Grok 4.6

Grok 4.6 underwent a longer supplemental training run than Grok 4.5, using:

  • Curated model-generated data for reasoning and advanced technical concepts
  • High-quality engineering data
  • An improved optimizer and training recipe

This provided the foundation for the supervised fine-tuning (SFT) and reinforcement learning (RL) stages that followed.

xAI then used Grok 4.5 to regenerate SFT trajectories across different reasoning efforts, agent harnesses, and domains such as STEM, software engineering, and knowledge work. Problematic traces were filtered using model-based checks. The resulting SFT checkpoint showed strong performance and improved behavior.

Grok 4.6 was trained on a wide range of agentic RL tasks, including:

  • Knowledge work
  • General coding
  • Kernel optimization
  • Web development
  • Computer-aided design (CAD)
  • More domain-specific environments
Long-running and interactive projects

xAI tested Grok 4.6 on projects designed to assess its range and ability to sustain work across many steps. The model can take a broad product idea and:

  • Research unfamiliar domains
  • Structure an application
  • Implement core interactions
  • Refine the result through several rounds of feedback

On longer trajectories, xAI observed more self-testing and verification, with the model checking its work before moving on.

Grok 4.6 also produced stronger first passes on visual and interactive projects than xAI typically saw with Grok 4.5. Given a concrete product idea, the model can establish an application’s structure and visual language in one pass, then continue iterating on the result.

Safety and safeguards

xAI says Grok 4.6’s safeguards have been improved and calibrated in line with the model’s capabilities.

The safety work covers legitimate use cases such as:

  • Vulnerability patching
  • Accelerating the engineering design cycle
  • Augmenting AI research

xAI’s safeguard evaluation included its widest-ever suite of pre-deployment testing for capabilities and safeguard calibration, along with extensive post-deployment and third-party testing.

Pricing and availability

Grok 4.6 is available today in Cursor and Grok Build. It is also available through the API and platforms including OpenRouter, Vercel, and Cloudflare.

Pricing starts at:

  • Input: $2 per million tokens
  • Output: $6 per million tokens
  • Fast variant: 2× the standard price

xAI is offering 2× included usage in Grok Build and Cursor for the first week.