cs.CV(2023-10-27)

📊 共 7 篇论文 | 🔗 1 篇有代码

🎯 兴趣领域导航

支柱九:具身大模型 (Embodied Foundation Models) (3) 支柱二:RL算法与架构 (RL & Architecture) (2 🔗1) 支柱三:空间感知与语义 (Perception & Semantics) (1) 支柱一:机器人控制 (Robot Control) (1)

🔬 支柱九:具身大模型 (Embodied Foundation Models) (3 篇)

#题目一句话要点标签🔗
1 3DCoMPaT$^{++}$: An improved Large-scale 3D Vision Dataset for Compositional Recognition 提出3DCoMPaT++以解决大规模3D视觉数据集的构建问题 multimodal
2 Image Clustering Conditioned on Text Criteria 提出基于文本标准的图像聚类方法以增强用户控制 large language model
3 Impressions: Understanding Visual Semiotics and Aesthetic Impact 提出Impressions数据集以解决图像美学与传播效果研究问题 multimodal

🔬 支柱二:RL算法与架构 (RL & Architecture) (2 篇)

#题目一句话要点标签🔗
4 Unsupervised Representation Learning for Diverse Deformable Shape Collections 提出无监督学习方法以处理多样变形形状集合 representation learning
5 ZeroNVS: Zero-Shot 360-Degree View Synthesis from a Single Image 提出ZeroNVS以解决复杂背景下的单图像新视角合成问题 distillation NeRF

🔬 支柱三:空间感知与语义 (Perception & Semantics) (1 篇)

#题目一句话要点标签🔗
6 Reconstructive Latent-Space Neural Radiance Fields for Efficient 3D Scene Representations 提出潜在空间神经辐射场以解决NeRF渲染速度慢的问题 NeRF neural radiance field

🔬 支柱一:机器人控制 (Robot Control) (1 篇)

#题目一句话要点标签🔗
7 One Style is All you Need to Generate a Video 提出基于风格的条件视频生成模型以提升视频质量 manipulation

⬅️ 返回 cs.CV 首页 · 🏠 返回主页