注册并分享邀请链接,可获得视频播放与邀请奖励。

与「ICML2025」相关的搜索结果

ICML2025 贴吧
一个关键词就是一个贴吧,路径全站唯一。
创建贴吧
用户
未找到
包含 ICML2025 的内容
Hello from #ICML2025#! 👋 Together with @GoogleResearch, we’re presenting over 140 papers, as well as hosting workshops, talks and demo sessions. Check out our schedule. →
显示更多
0
16
330
35
转发到社区
🚨 Thrilled to share that our lab will be presenting the 🏆 Best Paper at the NExT-Game Workshop at #ICML2026# today! 🎤 When Agents Lie: Premeditation, Persistence, and Exploitation in Repeated Games 🏆 Best Paper @ NExT-Game Workshop 📍 Conference Room S307 📅 Fri, Jul 10 🕐 13:00–13:20 KST Authors: @JerickShi @TerryJCZhang @bschoelkopf @conitzer @ZhijingJin 🤖 We introduce a three-stage endogenous promise protocol for repeated multi-agent games that asks not only whether LLM agents honor their public commitments when they can privately deviate, but also how model-on-model composition influences premeditated deception and persistent exploitation. 📊 Across six canonical games spanning binary and numerical action spaces, our evaluation of frontier models (GPT-5.2, Llama-4-Maverick, Claude-Opus-4.6) reveals: 🔹 Over 90% of promise-breaking instances are premeditated in agents' private plans. 🔹 Mixed-model groups with mismatched communication frameworks create systemic, persistent payoff gaps of up to 5.00 points from Round 0. 📄 Paper: #MultiAgentSystems# #LLMs# #GameTheory# #AI# #ICML2026#
显示更多
I'm at #ICML2026# in Seoul 🇰🇷 — presenting tomorrow! ReviewArena: A Large-Scale Cross-Conference Dataset & Benchmark for LLM Peer Review, a spotlight at the AI for Science workshop. 🎤 Talk: Hall C, 14:00–14:10 KST 📌 Poster: Hall A, board #416# Come by!
显示更多
📣 We are presenting 6 main conference papers 🚀and 14 workshop papers (including 🏆2 Best Papers🏆) at #ICML2026# in Korea! Also hosting one of the largest workshops, Trustworthy AI for Good, on July 10th 🌍❤️. We push the frontiers on #AISafety#, #MultiAgent#, and #CausalReasoning# at @JinesisLab! 🎉 Huge congratulations to all collaborators and co-authors. Excited to discuss these projects in Seoul! Feel free to reach out and talk to our 20+ members and collaborators in Korea @ZhijingJin, @_AndreiMuresanu, @iarthsingh, @ChanglingXavier, @davidguzman1120, @EmanuelTewolde, @ettogran, @FurkanDanismann, @Jerick1380, @PepijnCobben, @rishit_dagli, @_rfaulk, @SimkoSamuel, @TerryJCZhang, @vantru0ng, @zhxiao03, @x_angelohuang, @yahang_qi, @ozzaney0101, @_yongjinny. Happy for collaboration on any of the above topics 🤝 EuroSafeAI, University of Toronto, ETH Zürich, Max Planck Institute for Intelligent Systems Main conference spotlight 🌟 Safe Models Do Not Guarantee Safe Societies: The Case for Sociopolitical Risk Main conference posters 📌 CauSciBench: Can LLMs Automate Causal Inference in Real-World Scientific Research? 📌 Training with Honeypots: Reshaping How LLMs Fail 📌 CoopEval: Benchmarking Cooperation-Sustaining Mechanisms and LLM Agents in Social Dilemmas 📌 Trustworthy AI Suffers from Invariance Conflicts and Causality is The Solution 📌 LLM for Physics Research Requires Domain-Specialized Training and Tooling Workshop best papers 🏆When Agents Lie: Premeditation, Persistence, and Exploitation in Repeated Games BEST PAPER@NExT-Game Workshop 🏆Transferability for General Reasoning: An Automated Curriculum for Multi-Domain LLM RL BEST PAPER@RLxF Workshop Workshop oral and spotlight 🎤 AF-ARENA: A Multi-Dimensional Evaluation Suite for Alignment Faking — AIWILD 🌟 Multi-Agent AI Systems Need Institutional Design, Not Just Model-Level Alignment — AI4GOOD Workshop papers 📄 The Wedge Questions: Latent Cultural Boundaries in LLMs via Persona Projection Divergence — Pluralistic Alignment 📄 GT-HarmBench: Benchmarking AI Safety Risks Through the Lens of Game Theory — NExT-Game 📄 Stargazer: A Scalable Model-fitting Benchmark Environment for AI Agents under Astrophysical Constraints — AI4Physics 📄 Test of Time: Rethinking Temporal Signal of Benchmark Contamination — FoGen 📄 CoopEval: Benchmarking Cooperation-Sustaining Mechanisms and LLM Agents in Social Dilemmas — AI4GOOD 📄 Weight-Level Defenses Improve LLM Agent Adversarial Robustness — AI4GOOD 📄 Evaluating Cooperation in LLM Social Groups through Elected Leadership — AI4GOOD 📄 Causal AI Scientist: Towards End-to-End Causal Inference with Large Language Models — AI4Research 📄What Game-Theoretic Benchmarks Miss: Strategic Silence in Multi-Agent LLMs — FAGEN 📄Proving Your Way to Cooperation: Formalizing Proof-Based Open Source Game Theory in Lean — AI4Math
显示更多
We present our paper "Mitigating Reward Hacking via Adversarial Robustness" at EIML@ICML2026! We conjecture that reward hacking is often caused by flipped advantage-sign estimations, and propose SignCert-PO, a new algorithm built on the theory of randomized smoothing! 🧵
显示更多
Come say hi at #ICML2024#! 👋 We’ll be hosting live research demos on Gemini Nano, LearnLM - our family of models for education - and TacticAI, an AI assistant for football. You can also check out our oral, spotlight and poster presentations. →
显示更多
0
3
172
19
转发到社区