注册并分享邀请链接,可获得视频播放与邀请奖励。

与「QnA」相关的搜索结果

QnA 贴吧
一个关键词就是一个贴吧,路径全站唯一。
创建贴吧
用户
未找到
包含 QnA 的内容
Don’t open comments pls 🥶
How Buster Keaton did this insane stunt in 1924 😲😲😲
0
3
759
48
转发到社区
Grok Build is growing insanely fast, reaching 1.16 million monthly visits within just weeks after its early-beta launch Traffic jumped from around 460,000 visits in May to 1.16 million in June That’s growth of more than 150% in just one month Grok Build with Grok 4.5 ranked #1# on the SWE-Atlas-QnA benchmark with a score of 84, ahead of Claude Code with Fable 5 on the benchmark More developers are discovering Grok Build and realizing how capable Grok 4.5 actually is for real coding work Grok Build is quickly taking over the AI coding space
显示更多
0
61
422
44
转发到社区
NEW: Grok 4.5 now scores highest on the SWE-Atlas-QnA benchmark, edging Claude Fable 5 & GPT-5.6 Sol.
0
130
1.9K
91
转发到社区
Grok 4.5 with Grok Build has tied for the #1# spot on the SWE-Atlas-QnA benchmark, matching Codex GPT-5.6 with an impressive 84 score. Another strong milestone for xAI's rapidly advancing coding capabilities.
显示更多
0
25
129
28
转发到社区
Grok 4.5 with Grok Build just ranked #1# on the SWE-Atlas-QnA benchmark with a score of 84 That puts it level with GPT-5.6 (max) Codex and ahead of Claude Code Fable 5 (max), Opus 4.8 (max), and every other tested coding setup Grok Build is now the most powerful harness for agentic developer workflows
显示更多
0
29
170
21
转发到社区
Grok 4.5 is dominating the latest AI leaderboards Claims the #1# spot: • #1# on AutomationBench-AA • #1# on Terminal-Bench v2 • #1# on Harvey Legal Agent Benchmark • #1# on SWE Marathon • #1# on SWE-Atlas-QnA
显示更多
0
41
259
52
转发到社区
SpaceXAI’s Grok 4.5 scores 54 to place fourth on the Artificial Analysis Intelligence Index following only Fable 5, GPT-5.5, and Opus 4.8. It scores on par with GPT-5.5 in Codex on the Artificial Analysis Coding Agent Index in the Grok Build harness, at much lower cost Grok 4.5 improves 16 points over Grok 4.3 on the Intelligence Index, bringing SpaceXAI to the intelligence frontier behind only OpenAI and Anthropic, and outperforming all open weights models and notably Google’s Gemini models. Key standout areas of performance are agentic knowledge work and coding. Grok 4.5 in Grok Build scores 76 on the Artificial Analysis Coding Agent Index, on par with GPT-5.5 (xhigh) in Codex and just below Fable 5 (max) in Claude Code, and at a small fraction of the token usage and price. Congratulations to @SpaceXAI, @cursor_ai, and @elonmusk on the impressive release! Key Takeaways: ➤ Grok 4.5 performs very strongly on agentic tasks. Grok 4.5 ranks #4# on GDPval-AA v2 with an Elo of 1543, between Claude Opus 4.8 (1600) and GLM-5.2 (1513). It achieves the top score on 𝜏³-Banking of 33%, above 31% from GPT-5.5 (xhigh), and sits on the cost vs performance Pareto frontier across all three agentic evaluations in the Intelligence Index ➤ Grok 4.5 is one of the most cost efficient models to run for near-frontier intelligence. It costs $0.31 per task on the Artificial Analysis Intelligence Index and $2.59 per task on the Artificial Analysis Coding Agent Index within Grok Build ➤ Low cost for Grok 4.5 is driven by both low pricing and token efficiency. Grok 4.5 has a headline price over 60% lower than Claude Opus 4.8 and GPT-5.5, and used ~14k output tokens per Intelligence Index Task - over 60% lower than Opus 4.8. On the Coding Agent Index, Grok 4.5 stands out on the Pareto frontier of Coding Agent Index score vs. Total Tokens, using only 1.9M tokens for the Coding Agent Index while scoring 76 ➤ As a coding agent, Grok 4.5 in Grok Build is on par with GPT-5.5 and offers efficiency benefits: In our Artificial Intelligence Coding Agent Index that consists of DeepSWE, Terminal-Bench v2, and SWE-Atlas QnA, Grok 4.5 in Grok Build ranks third, on par with GPT-5.5 (Codex) and below Fable 5 (Claude Code). It is also very efficient in achieving this result: Grok 4.5 in Grok Build cost $2.49 per task while Fable 5 in Claude Code cost $11.80 and GPT-5.5 in Codex $5.07. This is driven by relatively low token pricing and the model using far fewer tokens than comparable models (1.9M average tokens used per task), significantly less than Fable 5 in Claude Code (7.2M) and GPT-5.5 in Codex (6.2M) Other model details: ➤ Context window of 500k tokens - a reduction from Grok 4.3’s 1M token context, but retaining configurable reasoning and vision input ➤ Pricing of $2/$6 per 1M tokens of input/output; cache hits are discounted by 75% to $0.5 per 1M tokens, and costs still double with long (>200k token) inputs ➤ As Elon Musk has disclosed, Grok 4.5 is 3x larger than its predecessor at 1.5T parameters
显示更多
0
39
442
65
转发到社区
アラサーにしてはイイ身体してると思う💁‍♀️
0
20
371
11
转发到社区
nudy version in comments (im serious)
0
1
74
66
转发到社区