跳到主要内容

agents

工具调用、规划、多智能体、记忆

12 篇,已拆 0 篇。回总览

2026

  • Agent-Computer Observation Interfaces Enable Dynamic Computer Use — paper=agent-computer-observation-interfaces-enable-dynamic@arXiv:2606.29472v1
  • RLVP: Penalize the Path, Reward the Outcome — paper=rlvp@arXiv:2607.07435v1
  • The Latent Bridge: A Continuous Slow-Fast Channel for Real-Time Game Agents — paper=the-latent-bridge-a-continuous-slow@arXiv:2606.24470v1
  • User as Code: Executable Memory for Personalized Agents — paper=user-as-code-executable-memory-for@arXiv:2606.16707v1
  • Whose Side Is Your Agent On? Multi-Party Principal Loyalty in LLM Agents — paper=whose-side-is-your-agent-on@arXiv:2606.30383v1

2025

  • DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models — paper=deepseek-v3-2@arXiv:2512.02556v1
  • DeepSeekMath-V2: Towards Self-Verifiable Mathematical Reasoning — paper=deepseekmath-v2@arXiv:2511.22570v1
  • MCP-Zero: Active Tool Discovery for Autonomous LLM Agents — paper=mcp-zero@arXiv:2506.01056v4

2024

  • Evaluating Very Long-Term Conversational Memory of LLM Agents — paper=evaluating-very-long-term-conversational-memory@arXiv:2402.17753v1

2023

  • MetaGPT: Meta Programming for A Multi-Agent Collaborative Framework — paper=metagpt@arXiv:2308.00352v7

2022

  • ReAct: Synergizing Reasoning and Acting in Language Models — paper=react@arXiv:2210.03629v3

2019

  • Mastering Atari, Go, Chess and Shogi by Planning with a Learned Model — paper=mastering-atari-go-chess-and-shogi@arXiv:1911.08265v2