| 1 |
Multimodal Spatiotemporal Atmospheric Data Assimilation with Latent Flow-matching |
提出基于潜在流匹配的多模态时空大气数据同化方法 |
flow matching spatiotemporal multimodal |
|
|
| 2 |
Agentic Reinforcement Learning with Observation-Calibrated Self-Distillation |
提出观察校准自蒸馏以解决稀疏奖励问题 |
reinforcement learning distillation privileged information |
✅ |
|
| 3 |
Helping Music Co-Creation Agents 'Listen' Well: Hierarchical Self-Supervised World Models for Understanding and Generation |
提出层次自监督世界模型以增强音乐共创代理的理解与生成能力 |
flow matching world model world models |
|
|
| 4 |
SpecRoll: Fast-Slow Verifier-Feedback Adaptation for Speculative Reinforcement Learning Rollouts |
提出SpecRoll以解决强化学习中生成效率瓶颈问题 |
reinforcement learning large language model |
|
|
| 5 |
NodeJEPA: Structure-Conditioned Latent Prediction for Node-Level Graph Self-Supervised Learning |
提出NodeJEPA以解决节点级图自监督学习的结构依赖问题 |
JEPA Joint-Embedding Predictive Architecture joint-embedding predictive architecture |
✅ |
|
| 6 |
Learning When to Stop: Prefix-Optimal Dynamic Diffusion Policies for Continuous Control |
提出Prefix-Optimal生成策略以解决动态扩散策略的计算瓶颈问题 |
diffusion policy |
|
|
| 7 |
Optimizing What Policies Learn From: Recoverability-aware Rollout Intervention Learning |
提出Recoverability-Aware Intervention Learning以优化策略学习过程 |
reinforcement learning large language model |
|
|
| 8 |
Consistency-Driven Co-Evolution for Self-Supervised Cross-Representation Learning |
提出CoCoEvolve以解决跨表示学习中的一致性问题 |
representation learning |
✅ |
|
| 9 |
Attention, Anomalies! Handling Attention Layers in Unsupervised Federated Outlier Detection |
提出新聚合方法以解决联邦学习中的异常检测问题 |
representation learning large language model |
|
|
| 10 |
Differentiating Through Dual Prices: End-to-End Policy Learning Under Capacity Constraints |
提出双价格区分方法以解决资源分配中的容量约束问题 |
policy learning |
|
|
| 11 |
Discretization and Statistical Consistency of Functional Flow Matching |
提出功能流匹配的离散化与统计一致性方法以解决传感器极限问题 |
flow matching |
|
|
| 12 |
SPOT: Sparse Probing and Outcome Calibration for On-Policy Distillation |
提出SPOT以解决在策略蒸馏中的不确定性问题 |
distillation |
|
|
| 13 |
Not Every Divergence Should Be Suppressed: Counterfactual Recoverability in On-Policy Distillation |
提出反事实可恢复性方法以优化在线蒸馏训练 |
distillation |
|
|