What if your text prompts could generate 3D scenes that actually obey gravity and don’t clip through each other? Carnegie Mellon, HKU, HKUST, and Genesis AI present PAT3D. It combines vision-language models with a physics simulator to arrange objects into stable,
