<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>VHD-Play | Xinjie Shen</title><link>https://xinjie-shen.com/tag/vhd-play/</link><atom:link href="https://xinjie-shen.com/tag/vhd-play/index.xml" rel="self" type="application/rss+xml"/><description>VHD-Play</description><generator>Wowchemy (https://wowchemy.com)</generator><language>en-us</language><lastBuildDate>Sun, 27 Sep 2026 00:00:00 +0000</lastBuildDate><image><url>https://xinjie-shen.com/media/icon_hu646f7301b7fde7528ecdae8cec89fc29_9606_512x512_fill_lanczos_center_3.png</url><title>VHD-Play</title><link>https://xinjie-shen.com/tag/vhd-play/</link></image><item><title>A Chat History of Making a Research Figure</title><link>https://xinjie-shen.com/post/research-figures-with-ai/</link><pubDate>Sun, 27 Sep 2026 00:00:00 +0000</pubDate><guid>https://xinjie-shen.com/post/research-figures-with-ai/</guid><description>&lt;p>People often ask how I made the teaser figure for &lt;a href="https://arxiv.org/abs/2609.27321" target="_blank" rel="noopener">VHD-Play&lt;/a>. Here is the visual history of the chat, in order.&lt;/p>
&lt;h2 id="chat-history">Chat history&lt;/h2>
&lt;p>Prompts are reproduced exactly as written.&lt;/p>
&lt;h3 id="initial-brief">Initial brief&lt;/h3>
&lt;blockquote>
&lt;p>can you generate a figrue, like we aim to solve real world tasks but don&amp;rsquo;t have label. we believe there is hidden dynamics of any task, even not explicted can be written. to learn on this, so we generate from know dynamics to envs with verifiable reward, and then we can train and wish to generate to any real world then. this should be our banner or teaser and our main belief . and check this main paper, &lt;a href="https://arxiv.org/abs/2609.27321" target="_blank" rel="noopener">https://arxiv.org/abs/2609.27321&lt;/a>&lt;/p>
&lt;/blockquote>
&lt;h3 id="1-the-first-version-tried-to-explain-everything">1. The first version tried to explain everything&lt;/h3>
&lt;blockquote>
&lt;p>@Create image cool now create a clear figure to suggest this belief and 引诱人们读这个paper&lt;/p>
&lt;/blockquote>
&lt;p>
&lt;figure >
&lt;div class="d-flex justify-content-center">
&lt;div class="w-100" >&lt;img src="https://telegraph-image-2rp.pages.dev/file/AgACAgUAAyEGAASNhWkUAAPOarkgrthrToLXmmkQaAwzrz6mMgsAAjQWaxuDBMhVJRtWYjbk8fYBAAMCAAN3AAM9BA.png" alt="A dense first version containing the problem, belief, method, training pipeline, and applications." loading="lazy" data-zoomable />&lt;/div>
&lt;/div>&lt;/figure>
&lt;/p>
&lt;h3 id="2-simplify-the-structure">2. Simplify the structure&lt;/h3>
&lt;p>&lt;em>Regenerated from the same instruction—no new prompt was entered.&lt;/em>&lt;/p>
&lt;p>
&lt;figure >
&lt;div class="d-flex justify-content-center">
&lt;div class="w-100" >&lt;img src="https://telegraph-image-2rp.pages.dev/file/AgACAgUAAyEGAASNhWkUAAPNarkgrhhrUxhzrczNkJexCLnIDawAAjMWaxuDBMhVKBd5hOYtww0BAAMCAAN3AAM9BA.png" alt="A cleaner three-panel version contrasting real-world tasks and generated training worlds." loading="lazy" data-zoomable />&lt;/div>
&lt;/div>&lt;/figure>
&lt;/p>
&lt;h3 id="3-switch-to-a-hand-drawn-visual-language">3. Switch to a hand-drawn visual language&lt;/h3>
&lt;blockquote>
&lt;p>good direction, but if maybe more intutive? and use like hand or pencil style?&lt;/p>
&lt;/blockquote>
&lt;p>
&lt;figure >
&lt;div class="d-flex justify-content-center">
&lt;div class="w-100" >&lt;img src="https://telegraph-image-2rp.pages.dev/file/AgACAgUAAyEGAASNhWkUAAPSarkgthJdLdr_UIkRt0GcaN5y2XgAAjgWaxuDBMhVS-LJW-TEhlgBAAMCAAN3AAM9BA.png" alt="A four-stage hand-drawn version moving from real-world tasks to generated environments, training, and deployment." loading="lazy" data-zoomable />&lt;/div>
&lt;/div>&lt;/figure>
&lt;/p>
&lt;h3 id="4-reduce-the-story-to-two-sides">4. Reduce the story to two sides&lt;/h3>
&lt;blockquote>
&lt;p>very good, maybe better stress? right now it seemed more text , like we can still have the two sides sense?&lt;/p>
&lt;/blockquote>
&lt;p>
&lt;figure >
&lt;div class="d-flex justify-content-center">
&lt;div class="w-100" >&lt;img src="https://telegraph-image-2rp.pages.dev/file/AgACAgUAAyEGAASNhWkUAAPParkgrv3ekKuhlJUFG1vV4o4r0U4AAjUWaxuDBMhVsuISHsnB4GQBAAMCAAN3AAM9BA.png" alt="A two-sided version connecting generated environments with known dynamics to real-world tasks with unknown dynamics." loading="lazy" data-zoomable />&lt;/div>
&lt;/div>&lt;/figure>
&lt;/p>
&lt;h3 id="5-make-known-dynamics-concrete">5. Make “known dynamics” concrete&lt;/h3>
&lt;blockquote>
&lt;p>我觉得这个已知的dynamics并不好懂，考虑试试看把他画得一眼能看出来是opreation research？&lt;/p>
&lt;/blockquote>
&lt;p>
&lt;figure >
&lt;div class="d-flex justify-content-center">
&lt;div class="w-100" >&lt;img src="https://telegraph-image-2rp.pages.dev/file/AgACAgUAAyEGAASNhWkUAAPMarkgrlnSJ8-cK1y9t7rRrLcYhmkAAjIWaxuDBMhV7PPPyBqDEoEBAAMCAAN3AAM9BA.png" alt="A revision showing route planning, scheduling, and resource allocation as concrete sources of training environments." loading="lazy" data-zoomable />&lt;/div>
&lt;/div>&lt;/figure>
&lt;/p>
&lt;h3 id="6-show-that-worlds-come-from-a-mathematical-model">6. Show that worlds come from a mathematical model&lt;/h3>
&lt;blockquote>
&lt;p>but we can&amp;rsquo;t tell we generate worlds vfrom known dynmaics or math?&lt;/p>
&lt;/blockquote>
&lt;p>
&lt;figure >
&lt;div class="d-flex justify-content-center">
&lt;div class="w-100" >&lt;img src="https://telegraph-image-2rp.pages.dev/file/AgACAgUAAyEGAASNhWkUAAPRarkgti1E5icjtBtf6eFkk8xpos0AAjcWaxuDBMhVBWNHOkHjY_EBAAMCAAN3AAM9BA.png" alt="A revision in which a known mathematical model is solved or simulated and rendered into training worlds." loading="lazy" data-zoomable />&lt;/div>
&lt;/div>&lt;/figure>
&lt;/p>
&lt;h3 id="7-make-the-causal-relationship-explicit">7. Make the causal relationship explicit&lt;/h3>
&lt;blockquote>
&lt;p>math mode;l is the dynmaics of genrated agentic env&lt;/p>
&lt;/blockquote>
&lt;p>
&lt;figure >
&lt;div class="d-flex justify-content-center">
&lt;div class="w-100" >&lt;img src="https://telegraph-image-2rp.pages.dev/file/AgACAgUAAyEGAASNhWkUAAPQarkgr5i3gtgP9GHXrBCB_3PQLCYAAjYWaxuDBMhVI1XsR2nAFeMBAAMCAAN3AAM9BA.png" alt="A numbered revision showing that the mathematical model defines the generated environment dynamics." loading="lazy" data-zoomable />&lt;/div>
&lt;/div>&lt;/figure>
&lt;/p>
&lt;h3 id="8-remove-the-extra-scaffolding">8. Remove the extra scaffolding&lt;/h3>
&lt;blockquote>
&lt;p>this is very very good, exactly what we want. but just curisopu if we can be more consice without losing infor, for example don&amp;rsquo;t need to draw the 1.2,3 so clearly?&lt;/p>
&lt;/blockquote>
&lt;p>
&lt;figure >
&lt;div class="d-flex justify-content-center">
&lt;div class="w-100" >&lt;img src="https://telegraph-image-2rp.pages.dev/file/AgACAgUAAyEGAASNhWkUAAPLarkgiqjcsc8SSzLRpgjggGgGIhMAAjEWaxuDBMhVEAMi_AlFqcEBAAMCAAN3AAM9BA.png" alt="The final VHD-Play teaser figure." loading="lazy" data-zoomable />&lt;/div>
&lt;/div>&lt;/figure>
&lt;/p></description></item></channel></rss>