cs.CV(2023-10-28)
📊 共 7 篇论文
🎯 兴趣领域导航
支柱九:具身大模型 (Embodied Foundation Models) (4)
支柱二:RL算法与架构 (RL & Architecture) (2)
支柱七:动作重定向 (Motion Retargeting) (1)
🔬 支柱九:具身大模型 (Embodied Foundation Models) (4 篇)
| # | 题目 | 一句话要点 | 标签 | 🔗 | ⭐ |
|---|---|---|---|---|---|
| 1 | UniCat: Crafting a Stronger Fusion Baseline for Multimodal Re-Identification | 提出UniCat以解决多模态重识别中的融合问题 | multimodal | ||
| 2 | CityRefer: Geography-aware 3D Visual Grounding Dataset on City-scale Point Cloud Data | 提出CityRefer数据集以解决城市级3D视觉定位问题 | visual grounding | ||
| 3 | Foundation Models for Generalist Geospatial Artificial Intelligence | 提出基础模型以解决地理空间人工智能中的多任务挑战 | foundation model | ||
| 4 | One-shot Localization and Segmentation of Medical Images with Foundation Models | 提出基于基础模型的一次性医学图像定位与分割方法 | foundation model |
🔬 支柱二:RL算法与架构 (RL & Architecture) (2 篇)
| # | 题目 | 一句话要点 | 标签 | 🔗 | ⭐ |
|---|---|---|---|---|---|
| 5 | Feature Guided Masked Autoencoder for Self-supervised Learning in Remote Sensing | 提出特征引导的掩蔽自编码器以提升遥感图像的自监督学习效果 | masked autoencoder MAE | ||
| 6 | Patch-Wise Self-Supervised Visual Representation Learning: A Fine-Grained Approach | 提出基于补丁级自监督学习的视觉表征学习方法以提升图像分类性能 | representation learning distillation |
🔬 支柱七:动作重定向 (Motion Retargeting) (1 篇)
| # | 题目 | 一句话要点 | 标签 | 🔗 | ⭐ |
|---|---|---|---|---|---|
| 7 | Triplet Attention Transformer for Spatiotemporal Predictive Learning | 提出三重注意力变换器以解决时空预测学习问题 | human motion spatiotemporal |