SCAFFOLD: A Large-Scale Structured Dataset of Computer Science Research Figures with Diagram QA and Chain-of-Thought Reasoning Traces
arXiv cs.AIen
arXiv cs.AI
AI Global WirearXiv:2609.00018v1 Announce Type: new Abstract: Computer science papers rely heavily on diagrams: architecture drawings, system flowcharts, and pipeline schematics that often carry more information than the text around them. There is currently no public dataset that pairs this specific kind of figure with captions, context, questions, answers, and step-by-step reasoning, which is exactly what is needed to train a vision-language model to understand them. We present \textbf{SCAFFOLD}\footnote{https://github.com/theranjitraut/scaffold}, a large-scale structured dataset of computer science research figures with diagram QA and Chain-of-Thought reasoning traces. This dataset consists of (image, c
This is a short summary published by AI Global Wire. The full article is owned and hosted by arXiv cs.AI — open it there to read it in full.
Read the full story at arXiv cs.AI- Forskning
- Bild
Related AI news
- UI-Venus-2 Technical ReportarXiv cs.AI · September 2, 2026
- When Prediction Error Is Not Enough: Evaluating Nuisance-Function Prediction for Causal EstimationarXiv cs.AI · September 2, 2026
- MiNER: Fine-Tuned Biomedical Natural Language Processing for Malaria Disease Entity Recognition in Clinical TextsarXiv cs.AI · September 2, 2026
- Asymmetries in Spontaneous and Instructed DeceptionarXiv cs.AI · September 2, 2026
- ReDeck: Step-Level Render-Grounded Refinement for Document-to-Slide GenerationarXiv cs.AI · September 2, 2026
- ConvDeck: Conversational Paper-to-Slide Generation via Stage-Specific User FeedbackarXiv cs.AI · September 2, 2026