mazpie

mazpie/genrl

[NeurIPS 2024] GenRL: Multimodal-foundation world models enable grounding language and video prompts into embodied domains, by turning them into sequences of latent world model states. Latent state sequences can be decoded using the decoder of the model, allowing visualization of the expected behavior, before training the agent to execute it.

⭐ 87 ⑂ 5 Python MIT · 2025-04-05推送
87
Watchers
0
贡献者
0
Commits
0
Releases
4
Open Issues
2025-04-05
最近推送
原文 中文
暂无 README