cs.CV(2026-08-12)
📊 共 34 篇论文 | 🔗 7 篇有代码
🎯 兴趣领域导航
支柱二:RL算法与架构 (RL & Architecture) (13 🔗2)
支柱九:具身大模型 (Embodied Foundation Models) (8 🔗2)
支柱三:空间感知与语义 (Perception & Semantics) (7 🔗2)
支柱七:动作重定向 (Motion Retargeting) (3)
支柱一:机器人控制 (Robot Control) (3 🔗1)
🔬 支柱二:RL算法与架构 (RL & Architecture) (13 篇)
🔬 支柱九:具身大模型 (Embodied Foundation Models) (8 篇)
🔬 支柱三:空间感知与语义 (Perception & Semantics) (7 篇)
🔬 支柱七:动作重定向 (Motion Retargeting) (3 篇)
| # | 题目 | 一句话要点 | 标签 | 🔗 | ⭐ |
|---|---|---|---|---|---|
| 29 | Autonomous Telerehabilitation via Skeletal Motion Prediction and Joint-Level Performance Assessment | 提出基于骨骼运动预测的自主远程康复系统以解决缺乏持续监督的问题 | human motion motion prediction | ||
| 30 | HSTGFormer: Hyper Spatial-Temporal Graph Transformer for 3D Human Pose Estimation | 提出HSTGFormer以解决3D人类姿态估计中的时空信息分离问题 | human motion | ||
| 31 | PolarSym: Polar Geometry-aware Attention for CAD Floorplan Parsing | 提出PolarSym以解决CAD平面解析中的几何对称性问题 | spatial relationship |
🔬 支柱一:机器人控制 (Robot Control) (3 篇)
| # | 题目 | 一句话要点 | 标签 | 🔗 | ⭐ |
|---|---|---|---|---|---|
| 32 | Motion-as-Prompt: Enhancing Motion Reasoning in Multimodal Large Language Models via Motion-Guided Cross-Frame Visual Prompting | 提出Motion-as-Prompt以解决多模态大语言模型的运动推理问题 | manipulation large language model multimodal | ✅ | |
| 33 | AVA-Encoder: Towards Agent-Native Video Representation Learning | 提出AVA-Encoder以解决代理智能视频表示学习问题 | manipulation representation learning | ||
| 34 | A Hybrid Framework of Vision Transformer and Gated Recurrent Unit for Detection of Mosquito Diseases | 提出混合框架结合视觉变换器与门控循环单元以检测蚊子疾病 | locomotion |