Page 1 - Showing 8 of 12 posts
View all posts by years →
-
从0构建大模型(1)走完大模型训练流程
11 min 中文 - Lecture 10-Embodied Reasoning and Test-time Scaling
具身推理与测试时扩展:从 Embodied Chain-of-Thought 到 GRPO、视觉推理和机器人规划
18 min 中文 - Lecture 11-Frontier and Open Problems
机器人学习前沿与开放问题:泛化、数据、具身智能、终身学习与研究方法
25 min 中文 -
Lecture 4&5 — Reinforcement Learning强化学习 I/II:从 Bellman 方程到 DQN、PPO 与 SAC
30 min 中文 - Lecture 6-Generative Models
生成模型在机器人学习中的作用:从 VAE、VQ-VAE 到 diffusion policy 与 flow matching
15 min 中文 - Lecture 7-Sequence Modeling and Transformers
序列建模、Transformer、VLM 与机器人动作 tokenization
16 min 中文 - Lecture 8-World Models
世界模型:从 action-conditioned prediction 到 Dreamer、video world model 与 JEPA
16 min 中文 - Lecture 9-Generalist Robot Policies
通用机器人策略:Open X-Embodiment、Octo、CrossFormer、VLA 和可扩展评估
16 min 中文