| 1 |
CM2: Multimodal Cultural Reasoning via an Integrated Multi-Agent Framework |
提出CM2框架以解决多模态文化推理问题 |
large language model multimodal |
|
|
| 2 |
GarmentWeaver: Schema-Aware Structured Synthesis for Multimodal Sewing Patterns |
提出GarmentWeaver以解决多模态缝纫图案生成问题 |
multimodal |
|
|
| 3 |
Cross-Regional Grapevine Cold Hardiness Prediction via Learned Multimodal Latent Representations |
提出跨区域葡萄藤耐寒性预测框架以解决数据稀缺问题 |
multimodal |
|
|
| 4 |
TuringLLM: Efficiently Scaling Foundation Models Toward Physical AI |
提出Turing-20B-A2B以提升物理AI应用的效率与性能 |
foundation model |
|
|
| 5 |
Dense Clinical Contrasts Enhance Medical Knowledge Updating in Large Language Models |
提出SEER-Bench以增强大语言模型的医学知识更新能力 |
large language model |
|
|
| 6 |
VIBE: Video Instruction-aligned Background music gEneration |
提出VIBE以解决视频到音乐生成中的语义控制问题 |
multimodal instruction following |
|
|
| 7 |
MNIST-PRO: MNIST is Back as a Partially Observable World for AI Agents |
提出MNIST-PRO以解决部分可观察环境中的感知问题 |
multimodal |
|
|
| 8 |
Autoregressive Mosaics: Probing 2D Spatial Reasoning in Text-Only Language Models |
提出自回归马赛克基准以探测文本模型的二维空间推理能力 |
large language model |
|
|
| 9 |
Automated Testing of LLM-Based Post Hoc Explainers Using Model Checking as an Oracle |
提出基于模型检查的自动化测试方法以验证LLM生成的解释 |
large language model |
|
|
| 10 |
Wrong Prediction, Right Answer: Recovering Evidence from Collapsed LLM Sequence Scores |
提出一种新方法以恢复大语言模型的推理能力 |
large language model |
|
|
| 11 |
HSRM: Hidden-State Reward Models for Test-Time Verification |
提出HSRM以解决大型语言模型验证效率问题 |
large language model |
|
|
| 12 |
On the Prospects of Dynamic LLM Conversations in Software Development |
通过动态干预提升开发者与LLM的交互质量 |
large language model |
|
|
| 13 |
ATLAS: Dual-Horizon Diagnostic Evaluation for Industrial Tool-Use Agents |
提出ATLAS框架以解决工业工具使用代理的评估问题 |
large language model |
|
|
| 14 |
Generative Retrieval for E-commerce: Jointly Learning Embedding and Codebook with Same Product Cluster |
提出联合学习嵌入与码本以解决电商检索精度问题 |
large language model |
|
|
| 15 |
Designing an Auditable LLM-Supported Workflow for Qualitative Thematic Analysis |
提出可审计的LLM支持定性主题分析工作流程 |
large language model |
|
|
| 16 |
Towards Cognitive Process-Aware Proactive Writing Support |
提出基于认知过程的主动写作支持以解决创作负担问题 |
large language model |
|
|
| 17 |
DERELAB: Probing Defeasible Reasoning and Confirmation Bias in LLMs with a Generative Benchmark |
提出DeReLab框架以研究大型语言模型中的可撤销推理与确认偏差 |
large language model |
|
|
| 18 |
Augmenting Human Performance with an XR Agent Learning from Online Behavior and BCI Evidence |
提出OLIVE框架以增强高压动态任务中的人类表现 |
foundation model |
|
|
| 19 |
Beyond Ranking Accuracy: Evaluating LLM-Cited Feature Rationales for Next Basket Repurchase Recommendation |
提出基于LLM的特征理由以改善下一篮子复购推荐 |
large language model |
|
|
| 20 |
Generating Workflow DAGs from Natural Language with Non-Reasoning LLMs |
提出神经符号分解方法以生成企业工作流DAG |
large language model |
|
|