Hints-In-Browser: Benchmarking Language Models for Programming Feedback Generation
Fuente:
arXiv
Salvato in:
| Autori principali: | Kotalwar, Nachiket, Gotovos, Alkis, Singla, Adish |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Inference-Time Personalized Alignment with a Few User Preference Queries
di: Pădurean, Victor-Alexandru, et al.
Pubblicazione: (2025)
di: Pădurean, Victor-Alexandru, et al.
Pubblicazione: (2025)
Prompt Programming: A Platform for Dialogue-based Computational Problem Solving with Generative AI Models
di: Pădurean, Victor-Alexandru, et al.
Pubblicazione: (2025)
di: Pădurean, Victor-Alexandru, et al.
Pubblicazione: (2025)
Synthesizing High-Quality Programming Tasks with LLM-based Expert and Student Agents
di: Nguyen, Manh Hung, et al.
Pubblicazione: (2025)
di: Nguyen, Manh Hung, et al.
Pubblicazione: (2025)
Corruption Robust Offline Reinforcement Learning with Human Feedback
di: Mandal, Debmalya, et al.
Pubblicazione: (2024)
di: Mandal, Debmalya, et al.
Pubblicazione: (2024)
Proximal Curriculum with Task Correlations for Deep Reinforcement Learning
di: Tzannetos, Georgios, et al.
Pubblicazione: (2024)
di: Tzannetos, Georgios, et al.
Pubblicazione: (2024)
Benchmarking Generative Models on Computational Thinking Tests in Elementary Visual Programming
di: Pădurean, Victor-Alexandru, et al.
Pubblicazione: (2024)
di: Pădurean, Victor-Alexandru, et al.
Pubblicazione: (2024)
Automating Human Tutor-Style Programming Feedback: Leveraging GPT-4 Tutor Model for Hint Generation and GPT-3.5 Student Model for Hint Validation
di: Phung, Tung, et al.
Pubblicazione: (2023)
di: Phung, Tung, et al.
Pubblicazione: (2023)
Neural Task Synthesis for Visual Programming
di: Pădurean, Victor-Alexandru, et al.
Pubblicazione: (2023)
di: Pădurean, Victor-Alexandru, et al.
Pubblicazione: (2023)
Towards Generalizable Agents in Text-Based Educational Environments: A Study of Integrating RL with LLMs
di: Radmehr, Bahar, et al.
Pubblicazione: (2024)
di: Radmehr, Bahar, et al.
Pubblicazione: (2024)
Learning Embeddings for Sequential Tasks Using Population of Agents
di: Mahajan, Mridul, et al.
Pubblicazione: (2023)
di: Mahajan, Mridul, et al.
Pubblicazione: (2023)
Humanizing Automated Programming Feedback: Fine-Tuning Generative Models with Student-Written Feedback
di: Pădurean, Victor-Alexandru, et al.
Pubblicazione: (2025)
di: Pădurean, Victor-Alexandru, et al.
Pubblicazione: (2025)
Program Synthesis Benchmark for Visual Programming in XLogoOnline Environment
di: Wen, Chao, et al.
Pubblicazione: (2024)
di: Wen, Chao, et al.
Pubblicazione: (2024)
Divergent-Convergent Thinking in Large Language Models for Creative Problem Generation
di: Nguyen, Manh Hung, et al.
Pubblicazione: (2025)
di: Nguyen, Manh Hung, et al.
Pubblicazione: (2025)
Evolutionary Discovery of Reinforcement Learning Algorithms via Large Language Models
di: Sygkounas, Alkis, et al.
Pubblicazione: (2026)
di: Sygkounas, Alkis, et al.
Pubblicazione: (2026)
World Models with Hints of Large Language Models for Goal Achieving
di: Liu, Zeyuan, et al.
Pubblicazione: (2024)
di: Liu, Zeyuan, et al.
Pubblicazione: (2024)
Exploring the Impact of Quizzes Interleaved with Write-Code Tasks in Elementary-Level Visual Programming
di: Ghosh, Ahana, et al.
Pubblicazione: (2024)
di: Ghosh, Ahana, et al.
Pubblicazione: (2024)
The Right Kind of Help: Evaluating the Effectiveness of Intervention Methods in Elementary-Level Visual Programming
di: Ghosh, Ahana, et al.
Pubblicazione: (2025)
di: Ghosh, Ahana, et al.
Pubblicazione: (2025)
Large Language Models for In-Context Student Modeling: Synthesizing Student's Behavior in Visual Programming
di: Nguyen, Manh Hung, et al.
Pubblicazione: (2023)
di: Nguyen, Manh Hung, et al.
Pubblicazione: (2023)
Self-Hinting Language Models Enhance Reinforcement Learning
di: Liao, Baohao, et al.
Pubblicazione: (2026)
di: Liao, Baohao, et al.
Pubblicazione: (2026)
On Characterizations for Language Generation: Interplay of Hallucinations, Breadth, and Stability
di: Kalavasis, Alkis, et al.
Pubblicazione: (2024)
di: Kalavasis, Alkis, et al.
Pubblicazione: (2024)
Assessing Large Language Models for Automated Feedback Generation in Learning Programming Problem Solving
di: Silva, Priscylla, et al.
Pubblicazione: (2025)
di: Silva, Priscylla, et al.
Pubblicazione: (2025)
On the Limits of Language Generation: Trade-Offs Between Hallucination and Mode Collapse
di: Kalavasis, Alkis, et al.
Pubblicazione: (2024)
di: Kalavasis, Alkis, et al.
Pubblicazione: (2024)
Benchmarking the Robustness of Agentic Systems to Adversarially-Induced Harms
di: Nöther, Jonathan, et al.
Pubblicazione: (2025)
di: Nöther, Jonathan, et al.
Pubblicazione: (2025)
StepHint: Multi-level Stepwise Hints Enhance Reinforcement Learning to Reason
di: Zhang, Kaiyi, et al.
Pubblicazione: (2025)
di: Zhang, Kaiyi, et al.
Pubblicazione: (2025)
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints
di: Nöther, Jonathan, et al.
Pubblicazione: (2025)
di: Nöther, Jonathan, et al.
Pubblicazione: (2025)
Learning Neuro-symbolic Programs for Language Guided Robot Manipulation
di: Kalithasan, Namasivayam, et al.
Pubblicazione: (2022)
di: Kalithasan, Namasivayam, et al.
Pubblicazione: (2022)
Enhanced Conditional Generation of Double Perovskite by Knowledge-Guided Language Model Feedback
di: Lee, Inhyo, et al.
Pubblicazione: (2025)
di: Lee, Inhyo, et al.
Pubblicazione: (2025)
Learning to Hint for Reinforcement Learning
di: Xia, Yu, et al.
Pubblicazione: (2026)
di: Xia, Yu, et al.
Pubblicazione: (2026)
WebLLM: A High-Performance In-Browser LLM Inference Engine
di: Ruan, Charlie F., et al.
Pubblicazione: (2024)
di: Ruan, Charlie F., et al.
Pubblicazione: (2024)
Mitigating Distribution Sharpening in Math RLVR via Distribution-Aligned Hint Synthesis and Backward Hint Annealing
di: Xie, Pei-Xi, et al.
Pubblicazione: (2026)
di: Xie, Pei-Xi, et al.
Pubblicazione: (2026)
Efficiently Aligning Language Models with Online Natural Language Feedback
di: Ye, Christine, et al.
Pubblicazione: (2026)
di: Ye, Christine, et al.
Pubblicazione: (2026)
BrowserArena: Evaluating LLM Agents on Real-World Web Navigation Tasks
di: Anupam, Sagnik, et al.
Pubblicazione: (2025)
di: Anupam, Sagnik, et al.
Pubblicazione: (2025)
Pathology-Aware Multi-View Contrastive Learning for Patient-Independent ECG Reconstruction
di: Youssef, Youssef, et al.
Pubblicazione: (2026)
di: Youssef, Youssef, et al.
Pubblicazione: (2026)
The BrowserGym Ecosystem for Web Agent Research
di: De Chezelles, Thibault Le Sellier, et al.
Pubblicazione: (2024)
di: De Chezelles, Thibault Le Sellier, et al.
Pubblicazione: (2024)
Parameterized Argumentation-based Reasoning Tasks for Benchmarking Generative Language Models
di: Steging, Cor, et al.
Pubblicazione: (2025)
di: Steging, Cor, et al.
Pubblicazione: (2025)
MoE-Pruner: Pruning Mixture-of-Experts Large Language Model using the Hints from Its Router
di: Xie, Yanyue, et al.
Pubblicazione: (2024)
di: Xie, Yanyue, et al.
Pubblicazione: (2024)
FCoReBench: Can Large Language Models Solve Challenging First-Order Combinatorial Reasoning Problems?
di: Mittal, Chinmay, et al.
Pubblicazione: (2024)
di: Mittal, Chinmay, et al.
Pubblicazione: (2024)
FRAUD-RLA: A new reinforcement learning adversarial attack against credit card fraud detection
di: Lunghi, Daniele, et al.
Pubblicazione: (2025)
di: Lunghi, Daniele, et al.
Pubblicazione: (2025)
Optimal Learners for Realizable Regression: PAC Learning and Online Learning
di: Attias, Idan, et al.
Pubblicazione: (2023)
di: Attias, Idan, et al.
Pubblicazione: (2023)
Reflective Preference Optimization (RPO): Enhancing On-Policy Alignment via Hint-Guided Reflection
di: Zhao, Zihui, et al.
Pubblicazione: (2025)
di: Zhao, Zihui, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Inference-Time Personalized Alignment with a Few User Preference Queries
di: Pădurean, Victor-Alexandru, et al.
Pubblicazione: (2025) -
Prompt Programming: A Platform for Dialogue-based Computational Problem Solving with Generative AI Models
di: Pădurean, Victor-Alexandru, et al.
Pubblicazione: (2025) -
Synthesizing High-Quality Programming Tasks with LLM-based Expert and Student Agents
di: Nguyen, Manh Hung, et al.
Pubblicazione: (2025) -
Corruption Robust Offline Reinforcement Learning with Human Feedback
di: Mandal, Debmalya, et al.
Pubblicazione: (2024) -
Proximal Curriculum with Task Correlations for Deep Reinforcement Learning
di: Tzannetos, Georgios, et al.
Pubblicazione: (2024)