| 1 |
Rule-Compliant Visual Spatial Planning for Multimodal Large Language Models |
提出RuleMaze以解决多模态大语言模型的空间规划问题 |
large language model multimodal |
✅ |
|
| 2 |
Towards general embodied intelligence: integrating large language models, knowledge bases, and reasoning capabilities to build the next generation of AI agents |
提出整合大语言模型与知识库以推进通用体现智能 |
large language model multimodal |
|
|
| 3 |
EchoCoT: Extracting Hidden Chain-of-Thought from Large Reasoning Models |
提出EchoCoT以提取大型推理模型中的隐含思维链 |
chain-of-thought |
|
|
| 4 |
Frequency-Aware Continual Learning for Smart Contract Vulnerability Detection with Large Language Models |
提出频率感知的持续学习方法以解决智能合约漏洞检测问题 |
large language model |
|
|
| 5 |
MemTrapBench: Benchmarking Cognitive Traps in LLM Memory Use |
提出MemTrapBench以评估大语言模型中的认知陷阱 |
large language model |
|
|
| 6 |
Break It Down, Pass It On: Cross-Task Skill Transfer in LLM Agents |
提出跨任务技能转移方法以提升LLM代理的能力 |
large language model |
|
|
| 7 |
Optimal Skill Selection for LLM Agents with Provable Bicriteria Guarantees |
提出最佳前缀选择算法以优化LLM代理的技能选择 |
large language model |
|
|
| 8 |
Repo0: Design-Driven Zero-to-All Code Generation |
提出Repo0框架以解决零到全代码生成问题 |
large language model |
|
|
| 9 |
LLMs as Acquisition Policies for Finite-Pool Materials Optimization: A Controlled Study |
利用开放权重的大语言模型优化有限材料池的获取策略 |
large language model |
|
|
| 10 |
Volumetric Radiology AI in the Era of Multimodal Large Language Models |
提出多模态大语言模型以解决体积放射学AI的表示不匹配问题 |
large language model foundation model multimodal |
|
|
| 11 |
AEGIS: Preventing Cross-Domain Resource Abuse in MCP |
提出AEGIS以解决MCP中的跨域资源滥用问题 |
large language model multimodal |
|
|
| 12 |
Large Scale AI Grading of Handwritten Physics Assessments: Score Agreement and Olympiad Team Selection Outcomes |
基于GPT-5.5的手写物理评估AI评分系统提升评分一致性 |
multimodal |
|
|
| 13 |
FL-MAESTRO: Multi-Agent LLM Orchestration for Resource-Constrained Federated Learning |
提出FL-MAESTRO以解决资源受限的联邦学习中的决策问题 |
large language model |
✅ |
|
| 14 |
Terminal Agents: A Survey of AI Agents in Command-Line Environments |
提出终端智能体框架以整合命令行环境中的AI行为研究 |
large language model |
|
|
| 15 |
Peer-Voted LLM-Agent Stress Tests Find Feed-Induced Lexical Convergence but No Reliable Matched-Exposure Advantage for Distributed Sources |
提出PV-SST以评估LLM代理的词汇收敛性 |
large language model |
|
|
| 16 |
An LLM agent for end-to-end computational materials discovery |
提出MAESTRO以解决计算材料发现中的多尺度任务协调问题 |
large language model |
|
|