| 30 |
On the Geometry of Learned Representations in Event-Based Multi-Modal Egomotion Estimation |
提出多模态网络以改进事件驱动的自我运动估计 |
optical flow motion estimation |
|
|
| 31 |
On the Role of Depth in Surgical Vision Foundation Models: An Empirical Study of RGB-D Pre-training |
通过RGB-D预训练提升外科视觉基础模型性能 |
depth estimation scene understanding foundation model |
|
|
| 32 |
NoDrift3R: Raymap-Guided Coupling for Drift-Robust Unposed Feed-Forward 3D Reconstruction |
提出NoDrift3R以解决长序列重建中的位姿漂移问题 |
3D gaussian splatting 3DGS 3D reconstruction |
|
|
| 33 |
ImprovedVBGS: Real-time Continual Variational Bayes Gaussian Splatting |
提出ImprovedVBGS以解决实时持续重建问题 |
gaussian splatting splatting NeRF |
|
|
| 34 |
Training-Free Open-Vocabulary 3D Point-Cloud Segmentation on the Generalized Few-Shot Benchmark |
提出无训练的开放词汇3D点云分割方法以解决少样本问题 |
open-vocabulary open vocabulary |
|
|
| 35 |
Towards Domain-Generalized Open-Vocabulary Object Detection: A Progressive Domain-invariant Cross-modal Alignment Method |
提出渐进式领域不变跨模态对齐方法以解决开放词汇物体检测问题 |
open-vocabulary open vocabulary |
|
|
| 36 |
HybridSim: A Physics-Learning Hybrid Digital Twin for mmWave Human Sensing |
提出HybridSim以解决毫米波人类感知中的高保真信号模拟问题 |
3D gaussian splatting gaussian splatting splatting |
|
|
| 37 |
DPNeXt: A Lightweight Multi-Scale Feature Fusion Framework for Efficient ViT-Based Multi-Task Dense Prediction |
提出DPNeXt以解决多任务学习中的解码瓶颈问题 |
depth estimation scene understanding geometric consistency |
|
|
| 38 |
MotionForesight: Re-purposing Video Models for Future 3D Scene-Flow Prediction |
提出MotionForesight以解决未来3D场景流预测问题 |
scene flow human-object interaction |
|
|
| 39 |
Toward Semantic Communication for Real-time Mobile 3D Reconstruction |
提出语义通信框架以解决实时移动3D重建问题 |
3D reconstruction scene understanding |
|
|
| 40 |
Knowing the Self, Understanding the World: A Dual-Cognition Benchmark for UAV Spatio-temporal Reasoning with MLLMs |
提出UAV-DualCog以解决无人机双重认知能力不足问题 |
scene understanding large language model multimodal |
|
|
| 41 |
A Comprehensive Survey on Video Scene Parsing:Advances, Challenges, and Prospects |
综述视频场景解析的进展与挑战,提出未来研究方向 |
open-vocabulary open vocabulary foundation model |
|
|
| 42 |
PIXIE: A Zero-Shot texture-invariant 6D pose estimation framework for unseen objects with assembly defects |
提出PIXIE框架以解决无纹理物体的6D姿态估计问题 |
6D pose estimation |
|
|
| 43 |
ReCal3R: Reliability-Calibrated Learning Rates for Streaming 3D Reconstruction |
提出ReCal3R以解决流式3D重建中的可靠性问题 |
3D reconstruction |
|
|
| 44 |
CSS-BA: Gate-Guided Column Space Search for Bundle Adjustment |
提出Gate-Guided CSS-BA以解决低视差下的束调整问题 |
3D reconstruction |
|
|
| 45 |
Wan-Dancer: A Hierarchical Framework for Minute-scale Coherent Music-to-Dance Generation |
提出层次化框架以解决长时间音乐到舞蹈生成问题 |
optical flow |
|
|
| 46 |
QIRF Quantum-Inspired Non-Orthogonal Function-Space Compression for 3D Gaussian Splatting |
提出QIRF以解决3D Gaussian Splatting中的冗余压缩问题 |
3D gaussian splatting 3DGS gaussian splatting |
|
|
| 47 |
Exploration Matters for Escaping the Blur Trap in 3D Gaussian Splatting |
提出探索策略以解决3D高斯点云中的模糊陷阱问题 |
3D gaussian splatting 3DGS gaussian splatting |
✅ |
|
| 48 |
Robust Multimodal Dynamic Object Segmentation |
提出多模态动态物体分割框架以解决现有方法的不足 |
3D reconstruction scene reconstruction optical flow |
|
|
| 49 |
CaT-GS: Efficient 3DGS Rendering for Large Scale Scenes via Inter-frame Caching and Tile Scheduling |
提出CaT-GS以解决大规模场景渲染效率问题 |
3D gaussian splatting 3DGS gaussian splatting |
|
|
| 50 |
ReViV: Reconstructing the Viewer and the View in 4D from Monocular Egocentric Video |
提出ReViV框架以解决单目自我中心视频的4D重建问题 |
depth estimation egocentric multimodal |
✅ |
|
| 51 |
Plenoptic Condensation: A Novel Approach to Generalized Scene Reconstruction |
提出Plenoptic Condensation以解决场景重建精度不足问题 |
splatting scene reconstruction scene understanding |
✅ |
|
| 52 |
MuViSeg: Multi-View Segment Correspondences from Dense Geometry Priors |
提出MuViSeg以解决多视角图像分割对应问题 |
VGGT foundation model |
|
|
| 53 |
Locality-Aware Density Control for Efficient Gaussian-based Image Representation |
提出局部感知密度控制以解决高斯图像表示效率问题 |
gaussian splatting splatting |
✅ |
|
| 54 |
RayOcc: Occlusion-Aware Ray Occupancy Estimation via Gaussian Mixture Intensity |
提出RayOcc以解决相机3D语义占用预测中的遮挡问题 |
depth estimation |
|
|