01 / Generative AI
Reasoning-Augmented Diffusion
Generative AI / Structured Reasoning
I built a reasoning-aware text-to-image system that gives diffusion models an explicit intermediate representation before generation. Instead of treating a prompt as one fuzzy embedding, I designed an enriched scene graph that captures objects, counts, attributes, relationships, and reasoning constraints. I also built an image-editing chatbot where the scene graph acts as editable state: when a user asks to “add a window” or “make the couch green,” the graph updates directly and the image regenerates. This made image editing more programmable and controllable than repeatedly prompting a black-box diffusion model.

