Share this page:

FronTalk: Benchmarking Front-End Development as Conversational Code Generation with Multi-Modal Feedback

Xueqing Wu, Zihan Xue, Da Yin, Shuyan Zhou, Kai-Wei Chang, Nanyun Peng, and Yeming Wen, in COLM, 2026.

Code

Download the full text


Abstract


Bib Entry

@inproceedings{wu2026frontalk,
  title = {FronTalk: Benchmarking Front-End Development as Conversational Code Generation with Multi-Modal Feedback},
  author = {Wu, Xueqing and Xue, Zihan and Yin, Da and Zhou, Shuyan and Chang, Kai-Wei and Peng, Nanyun and Wen, Yeming},
  booktitle = {COLM},
  keyword_extra = {vlmodel},
  year = {2026}
}

Related Publications

  1. AutoSUIT Bench - Automated Security UnIt Test Benchmark for LLM Coding, ACL-Findings, 2026
  2. METAL: A Multi-Agent Framework for Chart Generation with Test-Time Scaling, ACL, 2025
  3. MQT-LLaVA: Matryoshka Query Transformer for Large Vision-Language Models, NeurIPS, 2024
  4. DACO: Towards Application-Driven and Comprehensive Data Analysis via Code Generation, NeurIPS (Datasets and Benchmarks Track), 2024
  5. VDebugger: Harnessing Execution Feedback for Debugging Visual Programs, EMNLP-Finding, 2024
  6. AVATAR: A Parallel Corpus for Java-Python Program Translation, ACL-Finding (short), 2023
  7. Unified Pre-training for Program Understanding and Generation, NAACL, 2021
  8. Retrieval Augmented Code Generation and Summarization, EMNLP-Finding, 2021