Papers with VTA
A Large-Scale Real-World Evaluation of an LLM-Based Virtual Teaching Assistant (2025.acl-industry)
Copied to clipboard
| Challenge: | Empirical studies on their effectiveness and acceptance in real-world classrooms are limited, leaving their practical impact uncertain. |
| Approach: | They develop an LLM-based virtual teaching assistant and deploy it in an introductory AI programming course with 477 graduate students. |
| Outcome: | The proposed system is tested in an introductory AI programming course with 477 graduate students. |
ALGOGEN: Tool-Generated Verifiable Traces for Reliable Algorithm Visualization (2026.findings-acl)
Copied to clipboard
| Challenge: | Recent LLM-based systems require simulation of algorithm flow and video rendering constraints. |
| Approach: | They propose a paradigm that decouples algorithm execution from rendering. |
| Outcome: | The proposed paradigm reduces execution success rates, element overlap, and inter-frame inconsistencies. |
Visual-Textual Alignment for Graph Inference in Visual Dialog (2020.coling-main)
Copied to clipboard
| Challenge: | Existing approaches to visual dialog do not understand semantic dependencies between visual and textual contents. |
| Approach: | They propose a Visual-Textual Alignment for Graph Inference network that makes up the lack of structural inference in visual dialog. |
| Outcome: | The proposed model outperforms existing models on a VisDial dataset. |
Bringing Pedagogy into Focus: Evaluating Virtual Teaching Assistants’ Question-Answering in Asynchronous Learning Environments (2025.findings-emnlp)
Copied to clipboard
| Challenge: | Existing assessments rely on surface-level metrics and lack sufficient grounding in educational theory . a new framework is proposed to evaluate VTAs in asynchronous learning environments . |
| Approach: | They propose a pedagogically-oriented evaluation framework tailored to asynchronous forum discussions . they construct classifiers using expert annotations of VTA responses on a diverse set of forum posts . |
| Outcome: | The proposed evaluation framework is rooted in learning sciences and tailored to asynchronous forum discussions. |