| 14 |
Self-Improving Large Language Models via Progressive Experience Evolution |
提出SPEE框架以解决自我提升语言模型的经验转化问题 |
reinforcement learning distillation large language model |
✅ |
|
| 15 |
Douyin Multimodal Embedding Model Technical Report |
提出Douyin多模态嵌入模型以解决高效检索与细粒度匹配问题 |
representation learning multimodal |
|
|
| 16 |
Disentangled Contrastive Learning for Zero-Shot Multilingual Dense Retrieval |
提出解耦对比学习以解决零样本多语言密集检索问题 |
representation learning contrastive learning zero-shot transfer |
|
|
| 17 |
IACM-RL: Intent-Aware Context Management and Reinforcement Learning for Complex Tool Invocation under Dynamic Intent Fluctuations |
提出IACM-RL以解决动态用户意图波动下的工具调用问题 |
reinforcement learning distillation |
|
|
| 18 |
CAVE: Competence-Aware Visual Boundary Evidence Alignment for Video Temporal Grounding |
提出CAVE以解决视频时间定位中的视觉证据对齐问题 |
reinforcement learning TAMP |
|
|
| 19 |
Look Ahead Before You Distill: Future Trajectory Validation of Teacher Guidance for Agentic On-Policy Distillation |
提出FutureBridge-OPD以解决多回合任务中的教师指导有效性问题 |
distillation |
✅ |
|
| 20 |
Learning What to Remember: Test-Time Training via Context Distillation |
提出测试时间上下文蒸馏方法以优化长上下文建模 |
distillation |
|
|
| 21 |
Bole: Efficient Tree Speculation for Hybrid-Attention Language Models |
提出Bole以解决混合注意力语言模型的树状推测效率问题 |
linear attention large language model |
|
|
| 22 |
Cross-Domain Hybrid OPD for Generalizable Search Agents |
提出跨域混合OPD框架以提升搜索代理的通用性 |
reinforcement learning distillation |
|
|