| 13 |
WilLaGS: Latent-Conditional 3D Appearance Fields for Robust Gaussian Splatting In-the-Wild |
提出WilLaGS以解决不受约束场景下的3D重建与外观合成问题 |
teacher-student 3D gaussian splatting 3DGS |
|
|
| 14 |
GAAT: Geometry-Aware Alignment Transformer for Multimodal UAV Perception |
提出GAAT以解决无人机多模态感知中的对齐问题 |
contrastive learning scene understanding foundation model |
|
|
| 15 |
uScenes: A Multimodal RGB and 3D Sonar Dataset for Underwater Robot Perception |
提出uScenes数据集以解决水下机器人感知问题 |
representation learning scene understanding multimodal |
✅ |
|
| 16 |
WALDO: One-Shot Exemplar-Conditioned Object Detection in Cluttered Scenes |
提出WALDO以解决复杂场景中的单次示例条件物体检测问题 |
world model world models JEPA |
|
|
| 17 |
Token-Budget Distillation: Transferring Full-Token Semantics to Compressed Video Vision-Language Models |
提出Token-Budget Distillation以解决视频VLM适应性高成本问题 |
teacher-student distillation |
|
|
| 18 |
Denoising-Aware Temporal Point Cloud Completion for 3D Crop Architecture Recovery and Phenotypic Trait Extraction |
提出Denoising-Aware方法以解决3D植物重建中的噪声与遮挡问题 |
Mamba MAE 3D reconstruction |
|
|
| 19 |
CommerceVibe: Learning to Design E-Commerce Creatives as Executable Visual Code via Dual-Feedback Reinforcement Learning |
提出CommerceVibe以解决电商创意生成中的结构化与可编辑性问题 |
reinforcement learning |
|
|
| 20 |
Relational Knowledge Distillation Brings DNN Representations Close Enough to Humans to Be Aligned Without Supervision |
提出关系知识蒸馏方法以实现DNN与人类表征的无监督对齐 |
distillation |
|
|
| 21 |
Dual-Stream Semantic Guidance with Prototype Anchor Calibration for Source-Fully-Free Adaptation of Vision-Language Models |
提出双流语义引导以解决视觉语言模型的源完全自由适应问题 |
teacher-student distillation |
✅ |
|