Papers by Ruilin Lv
ReLook: Vision-Grounded RL with a Multimodal LLM Critic for Agentic Web Coding (2026.acl-long)
Copied to clipboard
| Challenge: | Large Language Models (LLMs) excel at algorithmic code generation, but front-end development is lacking in visual fidelity and interaction. |
| Approach: | They propose an agentic, vision-grounded reinforcement learning framework that closes a loop by invoking a multimodal LLM as a tool. |
| Outcome: | The proposed framework outperforms baselines in front-end code generation. |