| 1 |
CHARM: A Multimodal Graph Foundation Model with Hierarchical Context Modeling for Zero-Shot Transfer |
提出CHARM以解决多模态图的零-shot迁移问题 |
large language model foundation model multimodal |
|
|
| 2 |
A Cost-Effective Multimodal LLM Reasoning Framework for Question Answering over Irregular Clinical Time Series |
提出ClinPRISM以解决不规则临床时间序列问答问题 |
large language model multimodal |
|
|
| 3 |
Large Language Model for Operations Research Formulation Selection in Multi-Warehouse Inventory Allocation |
提出基于大语言模型的多仓库库存分配优化方法 |
large language model |
|
|
| 4 |
KQFuzz: Knowledge-Guided Fuzzing for Quantum Libraries via Large Language Models |
提出KQFuzz以解决量子库模糊测试效率低下问题 |
large language model |
|
|
| 5 |
Beyond Epistemia: Epistemic Schizologia and Large Language Models as Techno-Semiotic Machines |
提出将大型语言模型视为技术符号机器以解决知识评估问题 |
large language model |
|
|
| 6 |
The Disruptive Impact of Large Language Models on Capture the Flag Competitions and the Path Toward Fair Play |
提出四组件保障框架以解决CTF竞赛中的公平性问题 |
large language model |
|
|
| 7 |
Salient Knowledge Pathways: Sparse Cross-Modal Routing for Efficient Knowledge-Intensive Multimodal Question Answering |
提出SKIP以解决知识密集型多模态问答的计算效率问题 |
multimodal |
✅ |
|
| 8 |
HANDBOOK.md: A Benchmark for Long-Context Agentic Instruction Following |
提出HANDBOOK.md基准以解决长上下文指令遵循问题 |
instruction following |
|
|
| 9 |
CADENCE: A Cardiac Atom Dictionary for Interpretable Neural Concept Extraction from ECG Foundation Models |
提出CADENCE框架以实现心电图模型的可解释性概念提取 |
foundation model |
|
|
| 10 |
Toward an Organizational Science of Multi-Agent LLM Systems: Decoupling Who, How, and Which Algorithm |
提出IMACS框架以解决多智能体LLM系统中的协作问题 |
large language model |
|
|
| 11 |
Laplace-PSN-IRT: Uncertainty Quantification for Neural Item Response Theory Models of LLM Benchmarks |
提出Laplace-PSN-IRT以解决神经IRT模型的不确定性量化问题 |
large language model |
|
|
| 12 |
Does Runtime Topology Context Improve LLM-Generated Kubernetes Security Patches? |
提出KuTIE以提升Kubernetes安全补丁生成的准确性 |
large language model |
|
|
| 13 |
Penelope: Localized Latent Recurrence for Efficient Structured Reasoning |
提出Penelope以解决复杂结构推理的效率问题 |
chain-of-thought |
|
|
| 14 |
Runtime Uncertainty Monitoring for LLM-Based Multi-Agent Systems Using Bayesian Networks |
提出基于贝叶斯网络的多智能体系统以解决不确定性监测问题 |
large language model |
|
|
| 15 |
How Do LLMs Read Bug Reports? An Empirical Study of Attention in LLMs for Automated Program Repair |
通过注意力模式分析提升LLM自动程序修复的成功率 |
large language model |
|
|
| 16 |
OmniQEC: discovering practical quantum error-correcting codes by an AI scientist |
提出OmniQEC以发现适用于现代量子处理器的量子纠错码 |
large language model |
|
|
| 17 |
HiSkill: Empowering LLM Agents with Hierarchical Skill Graphs |
提出HiSkill框架以解决LLM代理技能关系不足问题 |
large language model |
✅ |
|
| 18 |
Cognivia: A Cognitive Behavioral Therapy Copilot for Evidence-Based Mental Healthcare |
提出Cognivia以解决心理健康领域专业治疗师短缺问题 |
large language model |
✅ |
|
| 19 |
DecoEvo: Score-Decoupled Co-Evolution of Solver and Rubric-Generator Skills in Text Space |
提出DecoEvo以解决文本空间优化中的评估瓶颈问题 |
large language model |
|
|
| 20 |
OmniDelta: Skill-Driven Budget Allocation for Token Compression in OmniLLMs |
提出OmniDelta以解决OmniLLMs中的预算分配问题 |
large language model |
|
|
| 21 |
Why Public Service AI Governance Frameworks Risk Failing in the Age of General-Purpose AI: Lessons from Policing |
提出公共服务AI治理框架以应对通用AI挑战 |
large language model |
|
|
| 22 |
COVENANT: Natural-Language Workflow Compilation for Aligned Agent Execution |
提出COVENANT以解决大语言模型工作流对齐问题 |
large language model |
|
|
| 23 |
Specula: Scaling formal specifications for autonomous model checking of system code |
提出Specula以解决复杂系统代码的形式化规范生成问题 |
large language model |
✅ |
|
| 24 |
Hybrid Analysis for Secure MCP Tool Use in LLM Agents |
提出MTGuard以解决LLM代理工具使用安全问题 |
large language model |
|
|
| 25 |
RIDGE: An Autonomous Framework for Validation and Method Discovery in LLM-Generated Option Pricing |
提出RIDGE框架以验证和发现LLM生成的期权定价方法 |
large language model |
✅ |
|