| 1 |
GaussFusion: Towards Multimodal 3D Gaussian Pretraining |
提出GaussFusion以解决3D高斯表示的多模态预训练问题 |
representation learning MAE 3D gaussian splatting |
|
|
| 2 |
MoWorld: A Flash World Model |
提出MoWorld以解决高效实时世界模型构建问题 |
world model world models distillation |
|
|
| 3 |
What Images Cannot Say: Language-Guided Olfactory Representation Learning |
提出SCENT框架以解决视觉与嗅觉信号对齐问题 |
representation learning multimodal |
|
|
| 4 |
Generalized Synthetic Image Detection with Enhanced RGB-Noise Representation Learning |
提出RNSIDNet以解决合成图像检测的跨模型泛化问题 |
representation learning contrastive learning |
✅ |
|
| 5 |
Few-Medoids: An Embarrassingly Simple Coreset Selection Method for Few-Shot Knowledge Distillation |
提出Few-Medoids方法以解决少样本知识蒸馏中的核心集选择问题 |
teacher-student distillation |
✅ |
|
| 6 |
XRFormer: Multiscale Tokenization for XRF Representation Learning |
提出XRFormer以解决XRF光谱自动学习挑战 |
representation learning |
✅ |
|
| 7 |
Bridging Diffusion Pruning and Step Distillation with Teacher-Aligned Repair |
提出教师对齐修复方法以连接扩散剪枝与步骤蒸馏 |
distillation |
|
|
| 8 |
Straight-Path Flow Matching for Incomplete Multi-View Clustering |
提出流匹配框架以解决不完整多视图聚类问题 |
flow matching |
|
|
| 9 |
Progressive Reasoning with Primitive Correction for Compositional Zero-Shot Learning |
提出PRPC框架以解决组合零-shot学习中的错误传播问题 |
reinforcement learning chain-of-thought |
|
|
| 10 |
AlayaWorld: Long-Horizon and Playable Video World Generation |
提出AlayaWorld以解决游戏世界生成的高成本与低灵活性问题 |
world model world models |
|
|
| 11 |
MobileWan: Closing the Quality Gap for Mobile Video Diffusion |
提出MobileWan以解决移动视频生成质量不足问题 |
linear attention distillation |
✅ |
|