π Sources
Advances and Challenges in Foundation Agents
https://arxiv.org/pdf/2504.01990v1
Why do Multi-agent systems fail
https://arxiv.org/pdf/2503.13657
API vs. GUI Agents - Divergence and Convergence
https://arxiv.org/html/2503.11069v1
PaperBench - Evaluating AI's Ability to Replicate Research
https://arxiv.org/abs/2504.01848
MemInsight: Autonomous Memory Augmentation for LLM Agents
https://arxiv.org/pdf/2503.21760
BEARCUBS: A Benchmark for Computer-Using Web Agents
https://arxiv.org/pdf/2503.07919
AgentRxiv: Towards Collaborative Autonomous Research
https://arxiv.org/abs/2503.18102
PLAY2PROMPT: Zero-shot Tool Instruction Optimization for LLM Agents
https://arxiv.org/abs/2503.14432
Agents Play Thousands of 3D Video Games using PORTAL
https://arxiv.org/abs/2503.13356
Join my channel:
ππππππ
https://t.me/Artificial_Intelligence_Updates
Post #759
544

- β€ 1