跳转至

Sources

  1. Agent Skills. https://platform.claude.com/docs/en/agents-and-tools/agent-skills/overview.md
  2. Get started with Agent Skills in the API. https://platform.claude.com/docs/en/agents-and-tools/agent-skills/quickstart.md
  3. Skill authoring best practices. https://platform.claude.com/docs/en/agents-and-tools/agent-skills/best-practices.md
  4. Tool use with Claude. https://platform.claude.com/docs/en/agents-and-tools/tool-use/overview.md
  5. How to implement tool use. https://platform.claude.com/docs/en/agents-and-tools/tool-use/implement-tool-use.md
  6. Programmatic tool calling. https://platform.claude.com/docs/en/agents-and-tools/tool-use/programmatic-tool-calling.md
  7. Fine-grained tool streaming. https://platform.claude.com/docs/en/agents-and-tools/tool-use/fine-grained-tool-streaming.md
  8. Tool search tool. https://platform.claude.com/docs/en/agents-and-tools/tool-use/tool-search-tool.md
  9. Web search tool. https://platform.claude.com/docs/en/agents-and-tools/tool-use/web-search-tool.md
  10. Web fetch tool. https://platform.claude.com/docs/en/agents-and-tools/tool-use/web-fetch-tool.md
  11. Memory tool. https://platform.claude.com/docs/en/agents-and-tools/tool-use/memory-tool.md
  12. Code execution tool. https://platform.claude.com/docs/en/agents-and-tools/tool-use/code-execution-tool.md
  13. Bash tool. https://platform.claude.com/docs/en/agents-and-tools/tool-use/bash-tool.md
  14. Text editor tool. https://platform.claude.com/docs/en/agents-and-tools/tool-use/text-editor-tool.md
  15. Computer use tool. https://platform.claude.com/docs/en/agents-and-tools/tool-use/computer-use-tool.md
  16. MCP connector. https://platform.claude.com/docs/en/agents-and-tools/mcp-connector.md
  17. Remote MCP servers. https://platform.claude.com/docs/en/agents-and-tools/remote-mcp-servers.md
  18. What is the Model Context Protocol (MCP)? - Model Context Protocol. https://modelcontextprotocol.io/
  19. Agent SDK overview. https://platform.claude.com/docs/en/agent-sdk/overview.md
  20. Quickstart. https://platform.claude.com/docs/en/agent-sdk/quickstart.md
  21. Agent SDK reference - Python. https://platform.claude.com/docs/en/agent-sdk/python.md
  22. Agent SDK reference - TypeScript. https://platform.claude.com/docs/en/agent-sdk/typescript.md
  23. TypeScript SDK V2 interface (preview). https://platform.claude.com/docs/en/agent-sdk/typescript-v2-preview.md
  24. Agent Skills in the SDK. https://platform.claude.com/docs/en/agent-sdk/skills.md
  25. Subagents in the SDK. https://platform.claude.com/docs/en/agent-sdk/subagents.md
  26. Intercept and control agent behavior with hooks. https://platform.claude.com/docs/en/agent-sdk/hooks.md
  27. Configure permissions. https://platform.claude.com/docs/en/agent-sdk/permissions.md
  28. Custom Tools. https://platform.claude.com/docs/en/agent-sdk/custom-tools.md
  29. MCP in the SDK. https://platform.claude.com/docs/en/agent-sdk/mcp.md
  30. Structured outputs in the SDK. https://platform.claude.com/docs/en/agent-sdk/structured-outputs.md
  31. Tracking Costs and Usage. https://platform.claude.com/docs/en/agent-sdk/cost-tracking.md
  32. Rewind file changes with checkpointing. https://platform.claude.com/docs/en/agent-sdk/file-checkpointing.md
  33. Streaming Input. https://platform.claude.com/docs/en/agent-sdk/streaming-vs-single-mode.md
  34. Session Management. https://platform.claude.com/docs/en/agent-sdk/sessions.md
  35. Slash Commands in the SDK. https://platform.claude.com/docs/en/agent-sdk/slash-commands.md
  36. Plugins in the SDK. https://platform.claude.com/docs/en/agent-sdk/plugins.md
  37. Modifying system prompts. https://platform.claude.com/docs/en/agent-sdk/modifying-system-prompts.md
  38. Hosting the Agent SDK. https://platform.claude.com/docs/en/agent-sdk/hosting.md
  39. Securely deploying AI agents. https://platform.claude.com/docs/en/agent-sdk/secure-deployment.md
  40. Migrate to Claude Agent SDK. https://platform.claude.com/docs/en/agent-sdk/migration-guide.md
  41. Todo Lists. https://platform.claude.com/docs/en/agent-sdk/todo-tracking.md
  42. Building Effective AI Agents. https://www.anthropic.com/research/building-effective-agents
  43. Our framework for developing safe and trustworthy agents. https://www.anthropic.com/news/our-framework-for-developing-safe-and-trustworthy-agents
  44. Introducing the Model Context Protocol. https://www.anthropic.com/news/model-context-protocol
  45. Donating the Model Context Protocol and establishing the Agentic AI Foundation. https://www.anthropic.com/news/donating-the-model-context-protocol-and-establishing-of-the-agentic-ai-foundation
  46. Developing a computer use model. https://www.anthropic.com/news/developing-computer-use
  47. Introducing computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku. https://www.anthropic.com/news/3-5-models-and-computer-use
  48. Enabling Claude Code to work more autonomously. https://www.anthropic.com/news/enabling-claude-code-to-work-more-autonomously
  49. Claude Code and new admin controls for business plans. https://www.anthropic.com/news/claude-code-on-team-and-enterprise
  50. Anthropic acquires Bun as Claude Code reaches $1B milestone. https://www.anthropic.com/news/anthropic-acquires-bun-as-claude-code-reaches-usd1b-milestone
  51. Agentic Misalignment: How LLMs could be insider threats. https://www.anthropic.com/research/agentic-misalignment
  52. Simple probes can catch sleeper agents. https://www.anthropic.com/research/probes-catch-sleeper-agents
  53. Sleeper Agents: Training Deceptive LLMs that Persist Through Safety Training. https://www.anthropic.com/research/sleeper-agents-training-deceptive-llms-that-persist-through-safety-training
  54. Introducing advanced tool use on the Claude Developer Platform. https://www.anthropic.com/engineering/advanced-tool-use
  55. Building agents with the Claude Agent SDK. https://www.anthropic.com/engineering/building-agents-with-the-claude-agent-sdk
  56. Claude Code Best Practices. https://www.anthropic.com/engineering/claude-code-best-practices
  57. Making Claude Code more secure and autonomous with sandboxing. https://www.anthropic.com/engineering/claude-code-sandboxing
  58. The "think" tool: Enabling Claude to stop and think. https://www.anthropic.com/engineering/claude-think-tool
  59. Code execution with MCP: building more efficient AI agents. https://www.anthropic.com/engineering/code-execution-with-mcp
  60. Claude Desktop Extensions: One-click MCP server installation for Claude Desktop. https://www.anthropic.com/engineering/desktop-extensions
  61. Effective context engineering for AI agents. https://www.anthropic.com/engineering/effective-context-engineering-for-ai-agents
  62. Effective harnesses for long-running agents. https://www.anthropic.com/engineering/effective-harnesses-for-long-running-agents
  63. Equipping agents for the real world with Agent Skills. https://www.anthropic.com/engineering/equipping-agents-for-the-real-world-with-agent-skills
  64. How we built our multi-agent research system. https://www.anthropic.com/engineering/multi-agent-research-system
  65. Writing effective tools for AI agents - using AI agents. https://www.anthropic.com/engineering/writing-tools-for-agents
  66. MRKL Systems: A modular, neuro-symbolic architecture that combines large language models, external knowledge sources and discrete reasoning. https://arxiv.org/abs/2205.00445
  67. PAL: Program-aided Language Models. https://arxiv.org/abs/2211.10435
  68. ReAct: Synergizing Reasoning and Acting in Language Models. https://arxiv.org/abs/2210.03629
  69. Toolformer: Language Models Can Teach Themselves to Use Tools. https://arxiv.org/abs/2302.04761
  70. WebGPT: Browser-assisted question-answering with human feedback. https://arxiv.org/abs/2112.09332
  71. Do As I Can, Not As I Say: Grounding Language in Robotic Affordances. https://arxiv.org/abs/2204.01691
  72. Reflexion: Language Agents with Verbal Reinforcement Learning. https://arxiv.org/abs/2303.11366
  73. ReWOO: Decoupling Reasoning from Observations for Efficient Augmented Language Models. https://arxiv.org/abs/2305.18323
  74. MemGPT: Towards LLMs as Operating Systems. https://arxiv.org/abs/2310.08560
  75. Generative Agents: Interactive Simulacra of Human Behavior. https://arxiv.org/abs/2304.03442
  76. Voyager: An Open-Ended Embodied Agent with Large Language Models. https://arxiv.org/abs/2305.16291
  77. HuggingGPT: Solving AI Tasks with ChatGPT and its Friends in Hugging Face. https://arxiv.org/abs/2303.17580
  78. Gorilla: Large Language Model Connected with Massive APIs. https://arxiv.org/abs/2305.15334
  79. AutoGen: Enabling Next-Gen LLM Applications via Multi-Agent Conversation. https://arxiv.org/abs/2308.08155
  80. CAMEL: Communicative Agents for "Mind" Exploration of Large Language Model Society. https://arxiv.org/abs/2303.17760
  81. MetaGPT: Meta Programming for A Multi-Agent Collaborative Framework. https://arxiv.org/abs/2308.00352
  82. ChatDev: Communicative Agents for Software Development. https://arxiv.org/abs/2307.07924
  83. ToolLLM: Facilitating Large Language Models to Master 16000+ Real-world APIs. https://arxiv.org/abs/2307.16789
  84. AgentBench: Evaluating LLMs as Agents. https://arxiv.org/abs/2308.03688
  85. WebArena: A Realistic Web Environment for Building Autonomous Agents. https://arxiv.org/abs/2307.13854
  86. WebShop: Towards Scalable Real-World Web Interaction with Grounded Language Agents. https://arxiv.org/abs/2207.01206
  87. ALFWorld: Aligning Text and Embodied Environments for Interactive Learning. https://arxiv.org/abs/2010.03768
  88. SWE-bench: Can Language Models Resolve Real-World GitHub Issues?. https://arxiv.org/abs/2310.06770
  89. SWE-agent: Agent-Computer Interfaces Enable Automated Software Engineering. https://arxiv.org/abs/2405.15793
  90. Tree of Thoughts: Deliberate Problem Solving with Large Language Models. https://arxiv.org/abs/2305.10601
  91. Graph of Thoughts: Solving Elaborate Problems with Large Language Models. https://arxiv.org/abs/2308.09687
  92. StableToolBench: Towards Stable Large-Scale Benchmarking on Tool Learning of Large Language Models. https://arxiv.org/abs/2403.07714
  93. Mind2Web: Towards a Generalist Agent for the Web. https://arxiv.org/abs/2306.06070
  94. WebVoyager: Building an End-to-End Web Agent with Large Multimodal Models. https://arxiv.org/abs/2401.13919
  95. OSWorld: Benchmarking Multimodal Agents for Open-Ended Tasks in Real Computer Environments. https://arxiv.org/abs/2404.07972
  96. SWE-bench Goes Live!. https://arxiv.org/abs/2505.23419
  97. OSWorld-Human: Benchmarking the Efficiency of Computer-Use Agents. https://arxiv.org/abs/2506.16042
  98. Mobile-Agent-v3: Fundamental Agents for GUI Automation. https://arxiv.org/abs/2508.15144
  99. MCP-Bench: Benchmarking Tool-Using LLM Agents with Complex Real-World Tasks via MCP Servers. https://arxiv.org/abs/2508.20453
  100. MobileUse: A GUI Agent with Hierarchical Reflection for Autonomous Mobile Operation. https://arxiv.org/abs/2507.16853
  101. MagicGUI: A Foundational Mobile GUI Agent with Scalable Data Pipeline and Reinforcement Fine-tuning. https://arxiv.org/abs/2508.03700
  102. UI-Evol: Automatic Knowledge Evolving for Computer Use Agents. https://arxiv.org/abs/2505.21964
  103. Demystifying evals for AI agents. https://www.anthropic.com/engineering/demystifying-evals-for-ai-agents