| 1 |
An Enclosed Mode Is a Gauge Choice: Topology Relative to Reach in Certified Code World Models |
提出一种新方法以解决认证代码世界模型中的拓扑问题 |
world model world models |
|
|
| 2 |
SOMTab: Set-Order Mamba for Efficient Tabular In-Context Learning |
提出SOMTab以提高表格上下文学习的效率 |
Mamba foundation model |
|
|
| 3 |
HARTS: Efficient Agentic Reinforcement Learning for Hybrid-Attention Models over Arbitrary Rollout Trees |
提出HARTS以解决混合注意力模型中的高效强化学习问题 |
reinforcement learning linear attention |
|
|
| 4 |
REPLICANT: Learning Policies for Evading and Hardening Malware Detectors |
提出Replicant框架以增强恶意软件检测的鲁棒性 |
reinforcement learning deep reinforcement learning privileged information |
|
|
| 5 |
VISTA: Verifier-Informed Student-to-Teacher Adaptation for On-Policy Self-Distillation |
提出VISTA以解决现有自蒸馏方法的单向监督问题 |
distillation |
|
|
| 6 |
Beyond Flat Netlist: Hierarchical Graph Representation Learning for Scalable Analysis of Sequential Circuits |
提出DeepSeq3以解决工业电路分析中的时序动态建模问题 |
representation learning |
|
|
| 7 |
VICT: Verifier-Instrumented Credit Tracing for Long-Horizon LLM Agent Reinforcement Learning |
提出VICT以解决长时间跨度LLM代理的细粒度信用分配问题 |
reinforcement learning |
|
|
| 8 |
When Can Conditional Flow Matching Replace Pointwise Negative Log-Likelihood? |
提出条件流匹配替代点对点负对数似然的条件 |
flow matching |
|
|
| 9 |
PhyMamba: Physics-Modulated Mamba for Robust Battery Health Prognostics |
提出PhyMamba框架以解决电池健康预测中的挑战 |
Mamba |
|
|