cs.LG(2026-09-01)
📊 共 23 篇论文 | 🔗 3 篇有代码
🎯 兴趣领域导航
支柱九:具身大模型 (Embodied Foundation Models) (16 🔗1)
支柱二:RL算法与架构 (RL & Architecture) (6 🔗2)
支柱一:机器人控制 (Robot Control) (1)
🔬 支柱九:具身大模型 (Embodied Foundation Models) (16 篇)
🔬 支柱二:RL算法与架构 (RL & Architecture) (6 篇)
| # | 题目 | 一句话要点 | 标签 | 🔗 | ⭐ |
|---|---|---|---|---|---|
| 17 | REFACTOR-VLA: Unsupervised Library Learning of Typed Motor Programs | 提出REFACTOR-VLA以解决长时间任务中的行为抽象问题 | world model world models vision-language-action | ||
| 18 | NashDreamer: Model-Based Reinforcement Learning for Zero-Sum Imperfect-Information Games | 提出NashDreamer以解决零和不完全信息博弈中的模型学习问题 | reinforcement learning dreamer | ||
| 19 | CopyShield: A Cross-Level Benchmark of Copyright Defenses in LLMs | 提出CopyShield以解决大语言模型版权防御评估不一致问题 | DPO direct preference optimization large language model | ✅ | |
| 20 | Explore More, Drift Less: Outcome-Only Reinforcement Learning Can Suffice for Long-Horizon Interactive Agents | 提出CANOPY以解决长时间交互任务中的信号稀缺与策略漂移问题 | reinforcement learning | ✅ | |
| 21 | Subliminal Learning as Trait-Direction Drift: A Mechanism and Targeted Control under SFT Distillation | 提出特征方向漂移机制以控制隐性学习问题 | distillation | ||
| 22 | Conditional Flow Matching for ML-Based Inverse Design Problems | 提出条件流匹配以解决基于机器学习的逆向设计问题 | flow matching |
🔬 支柱一:机器人控制 (Robot Control) (1 篇)
| # | 题目 | 一句话要点 | 标签 | 🔗 | ⭐ |
|---|---|---|---|---|---|
| 23 | Provably Safe Sim-to-Real Transfer | 提出安全的模拟到现实转移方法以解决数据收集中的安全问题 | sim-to-real reinforcement learning |