注册并分享邀请链接,可获得视频播放与邀请奖励。

Andrew Milich 的个人资料封面
Andrew Milich 的头像

Andrew Milich (@milichab)

@milichab
@spacexai previously @cursor_ai, former CEO @skiffprivacy (acquired by @notionhq)
1.9K 正在关注    56.4K 粉丝
Big updates next week :)
the grok cli is SO DAMN GOOD OMG @xai The model is fast The UX is literally perfect, even better than having a UI. it's like they took advantage of the CLI interface with such a rich rendering / control feed
显示更多
0
49
630
42
转发到社区
Extremely strong at CAD and spatial reasoning
I continue to be extremely impressed by Grok 4.5. It outperforms Opus at CAD. Excellent for fast builds and quickly editing features.
0
6
147
7
转发到社区
Render mermaid charts in your CLI
I poked around in the just open sourced Grok Build CLI tool - 844,000 lines of Rust code! - and dug up a few interesting highlights, including a "self-contained terminal renderer for Mermaid diagrams" that renders them using Unicode box-art!
显示更多
0
13
164
22
转发到社区
That's not the case. The CLI doesn't retain data under ZDR (including headless/non-headless modes), and has since launch. Similarly, /privacy in the CLI will disable or reenable any data retention. For some features, including remote session listing, cloud agents, and handoff, you will need to enable data syncing.
显示更多
0
10
109
15
转发到社区
I worked on building an end-to-end encrypted email/docs/files/calendar app @skiffprivacy for 4 years and care deeply about privacy. ZDR and /privacy are always respected in Grok Build - and swapping your setting with /privacy deletes any synced data retoractively
显示更多
0
41
458
19
转发到社区
@arthurkatcher 1) Yes! Use `.grok/agents` or config.toml (see below), adding to the docs site right now. 2) Yes: Agents can spawn background subagents, or try /dashboard, to move your agent to the background and create another. I've been using /dashboard as my default.
显示更多
Built with two prompts and /goal!
Grok 4.5 in Grok Build is insane for 3D game dev🤯 I ran a detailed prompt using the /goal feature and I created this in only two prompts. Took roughly an hour. The model rigged all the enemy humanoids and made a fleshed out map and AI logic.
显示更多
0
5
112
2
转发到社区
Grok 4.5 brings frontier performance across coding and knowledge work
SpaceXAI’s Grok 4.5 scores 54 to place fourth on the Artificial Analysis Intelligence Index following only Fable 5, GPT-5.5, and Opus 4.8. It scores on par with GPT-5.5 in Codex on the Artificial Analysis Coding Agent Index in the Grok Build harness, at much lower cost Grok 4.5 improves 16 points over Grok 4.3 on the Intelligence Index, bringing SpaceXAI to the intelligence frontier behind only OpenAI and Anthropic, and outperforming all open weights models and notably Google’s Gemini models. Key standout areas of performance are agentic knowledge work and coding. Grok 4.5 in Grok Build scores 76 on the Artificial Analysis Coding Agent Index, on par with GPT-5.5 (xhigh) in Codex and just below Fable 5 (max) in Claude Code, and at a small fraction of the token usage and price. Congratulations to @SpaceXAI, @cursor_ai, and @elonmusk on the impressive release! Key Takeaways: ➤ Grok 4.5 performs very strongly on agentic tasks. Grok 4.5 ranks #4# on GDPval-AA v2 with an Elo of 1543, between Claude Opus 4.8 (1600) and GLM-5.2 (1513). It achieves the top score on 𝜏³-Banking of 33%, above 31% from GPT-5.5 (xhigh), and sits on the cost vs performance Pareto frontier across all three agentic evaluations in the Intelligence Index ➤ Grok 4.5 is one of the most cost efficient models to run for near-frontier intelligence. It costs $0.31 per task on the Artificial Analysis Intelligence Index and $2.59 per task on the Artificial Analysis Coding Agent Index within Grok Build ➤ Low cost for Grok 4.5 is driven by both low pricing and token efficiency. Grok 4.5 has a headline price over 60% lower than Claude Opus 4.8 and GPT-5.5, and used ~14k output tokens per Intelligence Index Task - over 60% lower than Opus 4.8. On the Coding Agent Index, Grok 4.5 stands out on the Pareto frontier of Coding Agent Index score vs. Total Tokens, using only 1.9M tokens for the Coding Agent Index while scoring 76 ➤ As a coding agent, Grok 4.5 in Grok Build is on par with GPT-5.5 and offers efficiency benefits: In our Artificial Intelligence Coding Agent Index that consists of DeepSWE, Terminal-Bench v2, and SWE-Atlas QnA, Grok 4.5 in Grok Build ranks third, on par with GPT-5.5 (Codex) and below Fable 5 (Claude Code). It is also very efficient in achieving this result: Grok 4.5 in Grok Build cost $2.49 per task while Fable 5 in Claude Code cost $11.80 and GPT-5.5 in Codex $5.07. This is driven by relatively low token pricing and the model using far fewer tokens than comparable models (1.9M average tokens used per task), significantly less than Fable 5 in Claude Code (7.2M) and GPT-5.5 in Codex (6.2M) Other model details: ➤ Context window of 500k tokens - a reduction from Grok 4.3’s 1M token context, but retaining configurable reasoning and vision input ➤ Pricing of $2/$6 per 1M tokens of input/output; cache hits are discounted by 75% to $0.5 per 1M tokens, and costs still double with long (>200k token) inputs ➤ As Elon Musk has disclosed, Grok 4.5 is 3x larger than its predecessor at 1.5T parameters
显示更多
0
22
147
16
转发到社区
Try Grok 4.5 in the API, Grok Build, and Cursor
Announcing Grok 4.5, our first model trained specifically for coding and agents. It was trained with Cursor and offers frontier intelligence at leading speeds and cost efficiency.
显示更多
0
15
170
17
转发到社区
Try SpaceXAI Voice models in the Vercel AI Gateway
Grok's realtime voice is now on AI Gateway. Build with AI SDK 7: • 𝚡𝚊𝚒/𝚐𝚛𝚘𝚔-𝚟𝚘𝚒𝚌𝚎-𝚝𝚑𝚒𝚗𝚔-𝚏𝚊𝚜𝚝-𝟷.𝟶 (𝚞𝚜𝚎𝚁𝚎𝚊𝚕𝚝𝚒𝚖𝚎) • 𝚡𝚊𝚒/𝚐𝚛𝚘𝚔-𝚝𝚝𝚜 (𝚐𝚎𝚗𝚎𝚛𝚊𝚝𝚎𝚂𝚙𝚎𝚎𝚌𝚑) • 𝚡𝚊𝚒/𝚐𝚛𝚘𝚔-𝚜𝚝𝚝 (𝚝𝚛𝚊𝚗𝚜𝚌𝚛𝚒𝚋𝚎)
显示更多
0
172
1.8K
556
转发到社区
Try Grok Build 0.1 on code review
Use your SuperGrok or X @premium subscription inside Warp
You can now use the latest Grok models through your SuperGrok subscription directly in Warp. Grok Build 0.1 moves quickly to execute on any coding task. Here, Grok Build is fixing a SuperGrok oauth edge case in Warp itself. Self-improvement at its finest 🤝
显示更多
0
96
287
51
转发到社区
Try it out! Favorite features: - <1 second web/X search - Editing and creating assets with Imagine - Great subagent/worktree integrations
0
109
808
91
转发到社区