| --- |
| license: apache-2.0 |
| language: |
| - en |
| - zh |
| --- |
| # Model Card for *MineDreamer* 🔥 |
|
|
| <!-- Provide a quick summary of what the model is/does. --> |
|
|
| [](https://arxiv.org/abs/2403.12037) |
|
|
| [](https://sites.google.com/view/minedreamer/main) |
|
|
| *MineDreamer* is an instructable embodied agent for simulated control and it is developed on top of recent advances in Multimodal Large Language Models (MLLMs) and diffusion models! |
|
|
|
|
|
|
| <p align="center"> |
| <img src="https://cdn-uploads.huggingface.co/production/uploads/63f08dc79cf89c9ed1bb89cd/S62I1Tn5qz5qJ3IkgMHH8.png" width=93%> |
| <p> |
| |
| *MineDreamer* can follow instructions steadily by employing a Chain-of-Imagination (CoI) mechanism to envision the step-by-step process of executing instructions and translating imaginations into more precise visual prompts tailored to the current state; subsequently, it generates keyboard-and-mouse actions to efficiently achieve these imaginations, |
|
|
|
|
| <p align="center"> |
| <img src="https://cdn-uploads.huggingface.co/production/uploads/63f08dc79cf89c9ed1bb89cd/LJxBMChCFng_RkXwUotfk.png" width=93%> |
| <p> |
|
|
|
|
| **This repo is used for hosting *MineDreamer*'s Q-Former checkpoints, which are the training stage 1 checkpoints for Imaginator.** |
|
|
| For more details or tutorials see https://github.com/Zhoues/MineDreamer. |