[AINews] SpaceXAI Grok 4.6 and Grok @Bot
- ID
- 13767
- Status
- summarized
- Published
- 13 Aug 2026, 9:53 AM
- Fetched
- 13 Aug 2026, 10:15 AM
- Provider
- Latent Space
- Category
- developer-ai
- Original URL
- https://www.latent.space/p/ainews-spacexai-grok-46-and-grok
- Source URL
- https://www.latent.space/feed
Summary
- Score
- 6.0
- Created
- 13 Aug 2026, 10:15 AM
- Tags
- Audience
- developersai_agent_usersai_ml_learners
What happened
xAI released Grok 4.6, a 1.5T parameter model focused on long-running agents and interactive/visual work, alongside Grok Bot (@bot), an early-beta AI teammate that signs into your tools and returns finished work. Artificial Analysis reports Grok 4.6 is cost-competitive on their private AA-Briefcase agentic knowledge work benchmark, ranking near the top while costing substantially less than leading rivals. The model was trained using Grok 4.5-regenerated SFT trajectories across reasoning, agent harnesses, STEM, software engineering, and knowledge work, plus agentic RL on tasks including kernel optimization, web development, and CAD.
Why it matters
If you're building or evaluating AI agent pipelines, Grok 4.6's lower cost-per-task on agentic benchmarks makes it worth benchmarking against your current model choice for long-horizon workloads. The Grok Bot beta also represents a new entrant in the AI teammate category alongside Claude Tag and Block's Buzz—worth watching if you're selecting a tool-integrated agent for your team, but it's early beta with no pricing or availability details yet.
Discussion angle
Compare Grok 4.6's cost-efficiency positioning on AA-Briefcase against what the audience actually pays for Claude/GPT in agent loops today—is 'second best but cheaper' enough to switch, or do agentic workflows need the top model?