We honestly didn't see blender demo would get so much attention. Actually it is for a test of overall model capabilites especially coding in long horizon task.
We started GLM-5.3-Flash in an empty folder and let it run for 12 hours without stepping in. By the end it built this blender scene, used around 100 million tokens.
GLM wrote python scripts through the blender cli and created .blend file. In the loop it kept checking rendered images and making changes, until the result looked good. No mcp was used for the whole run.
Before writing the prompt, we went through several rounds of discussion with GLM about how to shape the scene. I’ve talked in previous post that it's always better to fully ask and understand “what” before asking “how” to make GLM know you better(than you do). After discussing, we decided to use 16 fixed camera views, and something others that should go into prompt.
When we wrote the prompt, we tried to describe what a good result should look like in as much detail as possible. This seems more important than modeling instructions which we think GLM now can handle well inside.
The whole process might feel a little scary if you never tried it before. You can start by giving GLM our template and your own idea. Ask it help you with prompt.
We will include the full prompt for skyline bar demo at the end as reference. Hope you can enjoy and help improve. And please share more tips and thoughts if you have good ones:)
Prompt as reference: https://t.co/xlgS7e18az