| 1 |
SpatialQ: Understanding 3D Gaussian Splatting Scene Quality via Visual-based MLLM |
提出多模态质量评估框架以解决3D Gaussian Splatting场景质量评估问题 |
representation learning 3D gaussian splatting 3DGS |
|
|
| 2 |
JEPADepth: Masked Predictive Representation Learning for Self-Supervised Monocular Depth Estimation |
提出JEPADepth以解决自监督单目深度估计问题 |
JEPA Joint-Embedding Predictive Architecture joint-embedding predictive architecture |
|
|
| 3 |
StatePlay: State-Aware Game World Models for Mechanics-Consistent Generation |
提出StatePlay以解决游戏世界模型中状态一致性问题 |
world model world models |
|
|
| 4 |
Veritas++: Value-aware On-Policy Distillation for Perception-Enhanced AIGI Detection |
提出Veritas++以解决AIGI检测中的感知瓶颈问题 |
distillation large language model |
✅ |
|
| 5 |
SCALPEL: Semantic Cross-modal Alignment via LLM-Powered Encoder Learning for Medical Vision-Language Representation |
提出SCALPEL以解决医学多模态表示学习中的对齐问题 |
representation learning large language model multimodal |
|
|
| 6 |
R-SLPR: Region-based Small-to-Large Point-cloud Registration with Contrastive Learning |
提出R-SLPR以解决小规模点云与大规模点云配准问题 |
MAE contrastive learning |
|
|
| 7 |
SciFigAlign: Scoring Scientific Figures by Fine-tuned Alignment of Visuals with Manuscript Evidence |
提出SciFigAlign以解决科学图形评估问题 |
MAE multimodal |
|
|
| 8 |
DistillAlign: Coordinating Mode Covering and Mode Seeking in Autoregressive Video Distillation |
提出DistillAlign以解决自回归视频蒸馏中的分布对齐问题 |
distillation |
|
|
| 9 |
Long-Tailed 3D Point Cloud Dataset Distillation |
提出长尾3D点云数据集蒸馏方法以解决数据不平衡问题 |
distillation |
|
|