cs.CV(2023-10-27)
📊 共 7 篇论文 | 🔗 1 篇有代码
🎯 兴趣领域导航
支柱九:具身大模型 (Embodied Foundation Models) (3)
支柱二:RL算法与架构 (RL & Architecture) (2 🔗1)
支柱三:空间感知与语义 (Perception & Semantics) (1)
支柱一:机器人控制 (Robot Control) (1)
🔬 支柱九:具身大模型 (Embodied Foundation Models) (3 篇)
| # | 题目 | 一句话要点 | 标签 | 🔗 | ⭐ |
|---|---|---|---|---|---|
| 1 | 3DCoMPaT$^{++}$: An improved Large-scale 3D Vision Dataset for Compositional Recognition | 提出3DCoMPaT++以解决大规模3D视觉数据集的构建问题 | multimodal | ||
| 2 | Image Clustering Conditioned on Text Criteria | 提出基于文本标准的图像聚类方法以增强用户控制 | large language model | ||
| 3 | Impressions: Understanding Visual Semiotics and Aesthetic Impact | 提出Impressions数据集以解决图像美学与传播效果研究问题 | multimodal |
🔬 支柱二:RL算法与架构 (RL & Architecture) (2 篇)
| # | 题目 | 一句话要点 | 标签 | 🔗 | ⭐ |
|---|---|---|---|---|---|
| 4 | Unsupervised Representation Learning for Diverse Deformable Shape Collections | 提出无监督学习方法以处理多样变形形状集合 | representation learning | ||
| 5 | ZeroNVS: Zero-Shot 360-Degree View Synthesis from a Single Image | 提出ZeroNVS以解决复杂背景下的单图像新视角合成问题 | distillation NeRF | ✅ |
🔬 支柱三:空间感知与语义 (Perception & Semantics) (1 篇)
| # | 题目 | 一句话要点 | 标签 | 🔗 | ⭐ |
|---|---|---|---|---|---|
| 6 | Reconstructive Latent-Space Neural Radiance Fields for Efficient 3D Scene Representations | 提出潜在空间神经辐射场以解决NeRF渲染速度慢的问题 | NeRF neural radiance field |
🔬 支柱一:机器人控制 (Robot Control) (1 篇)
| # | 题目 | 一句话要点 | 标签 | 🔗 | ⭐ |
|---|---|---|---|---|---|
| 7 | One Style is All you Need to Generate a Video | 提出基于风格的条件视频生成模型以提升视频质量 | manipulation |