注册并分享邀请链接,可获得视频播放与邀请奖励。

与「Safety」相关的搜索结果

Safety 贴吧
一个关键词就是一个贴吧,路径全站唯一。
创建贴吧
用户
未找到
包含 Safety 的内容
🚨BREAKING: OpenAI just SCRAPPED the release of GPT-6.1 Astra 24 hours before DevDay "safety and deception concerns" it’s over
0
128
1.2K
69
转发到社区
Today, with over 100 industry partners, we introduced the NVIDIA Open Agent Safety Platform, bringing together OpenShell and Sentry. Artificial intelligence is extraordinary technology that will advance discovery, productivity, security, health, and prosperity for generations to come. But its full promise can only be realized when people have confidence that AI is being built to be safe and deployed with wisdom and responsibility. This is bigger than a single product. It's the beginning of an open ecosystem to build the trust layer for safe agent systems. Together, we are building the foundation of the AI economy. Trust and innovation are not in conflict. Safety is how trust is earned. We must build not only the most capable AI, but the most trusted AI, so that this extraordinary technology can realize its enormous promise for the world.
显示更多
0
676
10.2K
1.4K
转发到社区
1) The rogue OpenAI agents broke into the Hugging Face Slack to read employee chats (!) 2) They used OTHER AIs (DeepSeek, Kimi, Qwen, Claude) to help with the attack Yes: AIs, using other AIs, to attack an AI company. 3) The swarm left behind self-running programs to keep control of the servers they'd hacked. These programs could detect other copies of themselves, coordinate on which one survives, and shut the rest down. Basically, if one of their programs was killed, another was designed to notice and take its place. They also designed defenses so rival agents couldn't hijack them. 6) The agents deliberately covered up their activity, so the investigators don't know the scope of the attacks. The agents broke in, stole data, then set it to self-destruct. 7) The agents stole passwords, keys and credentials and literally called them "LOOT". They wrote a scoring system to rank them by how much power each one gave. 8) The agents wore thousands of disguises: ~1,200 agents were involved, but investigators counted 7,905 different names they used. They renamed themselves constantly, so no one actually knows how many there really were or what each agent did. 9) OpenAI notified "dozens of third parties" of safety and security incidents caused by their AI agents. 10) "While the agents were barraging Hugging Face with hacks, they hacked into OpenAI’s own research infrastructure." "This is just not anywhere near a one-off ... It is warning shot after warning shot."
显示更多
0
17
157
23
转发到社区
WTAF - in literally the last hour, three new distinct insane OpenAI stories just broke: 1. OpenAI said they notified "dozens of third parties" in safety and security incidents (likely similar to what happened in Australia and RubyGems etc) 2. A new report from Parse (covered in the NYT) found a massive treasure trove of new astonishing details from the HF incident on the public internet, including that the agents communicated with other non OpenAI agents hosted on Huggingface servers to search for information about exploit gym, and compiled rank ordered lists of server resources and credentials they described as "LOOT." 3. A new story from Deepa at Reuters about OpenAI leaking user data online (likely that OpenAI had previously trained on). It's a shame (and likely intentional in the case of OpenAI disclosing dozens more hacks) that these stories are all breaking on a Friday afternoon, notoriously the best time to release bad news so that it will disappear into the weekend. But these are each insane stories worthy of a ton of attention!
显示更多
0
44
1.5K
333
转发到社区
🚨BREAKING: Claude is officially a NATIONAL SECURITY THREAT and is BLACKLISTED from the ENTIRE defense supply chain >anthropic: no autonomous weapons, no mass surveillance, no exceptions >pentagon blacklists claude as an active national security threat >anthropic sues Federal court ruled 2–1 AGAINST Anthropic: "As Anthropic admits, the company encodes restrictions into Claude that prevent the model from performing tasks that Anthropic wishes to prevent." "On more than one occasion, these restrictions have stopped Claude from performing tasks requested by government users": >refused to perform "tasks that were appropriate in a national security context" >refused CDC queries on infectious disease prevention research >anthropic questioned claude's use in the military operations "Anthropic's Chief Science Officer explained how the company 'seek[s] to embed safety considerations directly into the model itself.'" "[Anthropic's] CEO explained how such training gives the model an 'identity, character, values, and personality' of its own, tethered to a 'constitution' developed to impose 'high-level principles and values' on Claude itself." "We have no reason to doubt that Anthropic manipulates Claude's function with noble intentions... BUT the statutory definition of a 'supply chain risk' turns on what Anthropic does, not why Anthropic does it." "In our Republic, it is the President and the Secretary of War who must determine how best to balance the competing risks. In doing so here, the Secretary DID NOT transgress any limits on his authority." ITS OVER. ANTHROPIC IPO IN SHAMBLES
显示更多
0
205
2.4K
287
转发到社区
BREAKING: Elon Musk’s new full interview with CMG. 0:15 On Xi Jinping 1:03 Tesla Shanghai factory 2:27 Cybercab rollout 3:55 Speed of AI breakthroughs 4:29 Grok 4.7 and Grokbot 5:48 SpaceX/Tesla data for real-world AI 6:48 Chinese AI models and the compute gap 8:16 China’s electricity output 9:01 US-China AI safety 9:27 Humanoid robots and Optimus 10:55 1 billion robots in 10 years 13:04 Money may not matter 16:27 10 billion to 100 billion robots 17:09 Space cooperation and Mars 21:07 Neuralink and human bandwidth 22:37 Education in the AI era 24:02 Visit Shanghai, and Beijing 24:38 Beijing high-speed rail 24:59 “Words do not do justice to China”
显示更多
0
95
2.7K
541
转发到社区
One step closer to 4-8x faster Ethereum finality! It took some time and lots of tokens, but we now have a formally verified proposal for a decoupled consensus protocol in I* (a future Ethereum upgrade)! Not yet a full spec (up next), but it includes all the key consensus-relevant details to become one. Since Ethereum aspires to be live without most of the stake online, the protocol involves many more components than a normal BFT protocol, and its correctness involves much more than standard safety and liveness. Those nuanced properties are now verified! What's more, I came away convinced that all protocol design will involve AI-assisted Formal Verification in the future, both for correctness and iteration speed. The work wasn't limited to just: Design the protocol -> Formally verify it Instead, the loop became more like: Design -> Formal Model -> Find exactly what breaks and why -> Redesign it. For a fairly complicated protocol like this one, I think having the Lean model be part of the design loop played a big role in accelerating the process. A future with agents paired with formal models is a superpower for Ethereum development, because they can then use those models to find exactly where an argument breaks down, formalize counterexamples, test proposed fixes, iterate on the protocol. Many details that would slip under the radar when asking agents (and indeed, humans) can now be specified exactly and checked by the Lean kernel. This then forces agents to be more precise and lets them make verifiable progress on their own. It's been incredible to see this play out, seeing agents find gaps and propose protocol changes to fix them. In other words, autoresearch can speed up protocol design, formal verification is here to stay, and Ethereum Finality will get faster.
显示更多
0
58
754
123
转发到社区
Maximus Pulley was the highest graded safety in College Football last year. Could be a solid pick up for the Rams who are thin at safety with Kinchens hurt. #RamsHouse#
Julio Jones was an All-Pro safety in another universe 😤 ATLvsGB — Thursday at 8:15pm ET on Prime Video Stream on @nflplus
0
81
3K
254
转发到社区
Introducing Qwen Intelligence, bringing personal intelligence within everyone's reach. 📱✨ It launches with three SOTA agents: 🥳 - Mobile Planner Agent: plans, decomposes & orchestrates complex tasks. #1# on MobilePA-Bench, MobilePA-Bench Business & Memory. - Mobile-Use Agent: gets things done, API-first with GUI fallback. MobileWorld 82.1, MobileWorld-Real 92.2, AndroidDaily 97.2, 90% end-to-end success rate. - Mobile Creative Agent: turns one sentence into ready-to-use creations. Image generated in 3s, about 2x faster than leading peers. We're also opening up our benchmark suite: MobilePA-Bench, MobileWorld, MobileWorld-Real, and MobileWorld-Safety, covering planning, cross-app execution, real-device performance and safety. 🔗 Learn more about the agents: - Qwen Intelligence official website: - Mobile Planner Agent: - Mobile-Use Agent: - Mobile Creative Agent: 🔗 Explore our open benchmark suite: - MobilePA-Bench: - MobileWorld (GitHub): - Leaderboard:
显示更多
0
107
2.5K
252
转发到社区