cs.LG(2023-10-14)
📊 共 6 篇论文 | 🔗 1 篇有代码
🎯 兴趣领域导航
支柱二:RL算法与架构 (RL & Architecture) (3)
支柱九:具身大模型 (Embodied Foundation Models) (2 🔗1)
支柱一:机器人控制 (Robot Control) (1)
🔬 支柱二:RL算法与架构 (RL & Architecture) (3 篇)
| # | 题目 | 一句话要点 | 标签 | 🔗 | ⭐ |
|---|---|---|---|---|---|
| 1 | STORM: Efficient Stochastic Transformer based World Models for Reinforcement Learning | 提出STORM以提高强化学习中的世界模型效率 | reinforcement learning world model world models | ||
| 2 | A Blockchain-empowered Multi-Aggregator Federated Learning Architecture in Edge Computing with Deep Reinforcement Learning Optimization | 提出区块链增强的多聚合器联邦学习架构以解决边缘计算中的安全问题 | reinforcement learning deep reinforcement learning | ||
| 3 | Mirage: Model-Agnostic Graph Distillation for Graph Classification | 提出Mirage以解决图神经网络训练资源不足问题 | distillation |
🔬 支柱九:具身大模型 (Embodied Foundation Models) (2 篇)
| # | 题目 | 一句话要点 | 标签 | 🔗 | ⭐ |
|---|---|---|---|---|---|
| 4 | Multimodal Federated Learning in Healthcare: a Review | 综述多模态联邦学习在医疗领域的应用与挑战 | multimodal | ||
| 5 | DPZero: Private Fine-Tuning of Language Models without Backpropagation | 提出DPZero以解决大语言模型私有微调中的内存与隐私问题 | large language model | ✅ |
🔬 支柱一:机器人控制 (Robot Control) (1 篇)
| # | 题目 | 一句话要点 | 标签 | 🔗 | ⭐ |
|---|---|---|---|---|---|
| 6 | Reduced Policy Optimization for Continuous Control with Hard Constraints | 提出减少策略优化算法以解决连续控制中的硬约束问题 | manipulation reinforcement learning |