PECC: Problem Extraction and Coding Challenges
Fuente:
arXiv
Guardado en:
| Autores principales: | Haller, Patrick, Golde, Jonas, Akbik, Alan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Sample-Efficient Language Modeling with Linear Attention and Lightweight Enhancements
por: Haller, Patrick, et al.
Publicado: (2025)
por: Haller, Patrick, et al.
Publicado: (2025)
What Matters in Linearizing Language Models? A Comparative Study of Architecture, Scale, and Task Adaptation
por: Haller, Patrick, et al.
Publicado: (2025)
por: Haller, Patrick, et al.
Publicado: (2025)
Fabricator: An Open Source Toolkit for Generating Labeled Training Data with Teacher LLMs
por: Golde, Jonas, et al.
Publicado: (2023)
por: Golde, Jonas, et al.
Publicado: (2023)
PISA-Bench: The PISA Index as a Multilingual and Multimodal Metric for the Evaluation of Vision-Language Models
por: Haller, Patrick, et al.
Publicado: (2025)
por: Haller, Patrick, et al.
Publicado: (2025)
BabyHGRN: Exploring RNNs for Sample-Efficient Training of Language Models
por: Haller, Patrick, et al.
Publicado: (2024)
por: Haller, Patrick, et al.
Publicado: (2024)
What Matters When Building Universal Multilingual Named Entity Recognition Models?
por: Golde, Jonas, et al.
Publicado: (2026)
por: Golde, Jonas, et al.
Publicado: (2026)
FiNERweb: Datasets and Artifacts for Scalable Multilingual Named Entity Recognition
por: Golde, Jonas, et al.
Publicado: (2025)
por: Golde, Jonas, et al.
Publicado: (2025)
Repetition over Diversity: High-Signal Data Filtering for Sample-Efficient German Language Modeling
por: Aynetdinov, Ansar, et al.
Publicado: (2026)
por: Aynetdinov, Ansar, et al.
Publicado: (2026)
MastermindEval: A Simple But Scalable Reasoning Benchmark
por: Golde, Jonas, et al.
Publicado: (2025)
por: Golde, Jonas, et al.
Publicado: (2025)
Familiarity: Better Evaluation of Zero-Shot Named Entity Recognition by Quantifying Label Shifts in Synthetic Training Data
por: Golde, Jonas, et al.
Publicado: (2024)
por: Golde, Jonas, et al.
Publicado: (2024)
Large-Scale Label Interpretation Learning for Few-Shot Named Entity Recognition
por: Golde, Jonas, et al.
Publicado: (2024)
por: Golde, Jonas, et al.
Publicado: (2024)
Pre-Training Curriculum for Multi-Token Prediction in Language Models
por: Aynetdinov, Ansar, et al.
Publicado: (2025)
por: Aynetdinov, Ansar, et al.
Publicado: (2025)
Question Decomposition for Retrieval-Augmented Generation
por: Ammann, Paul J. L., et al.
Publicado: (2025)
por: Ammann, Paul J. L., et al.
Publicado: (2025)
NoiseBench: Benchmarking the Impact of Real Label Noise on Named Entity Recognition
por: Merdjanovska, Elena, et al.
Publicado: (2024)
por: Merdjanovska, Elena, et al.
Publicado: (2024)
From Data to Knowledge: Evaluating How Efficiently Language Models Learn Facts
por: Christoph, Daniel, et al.
Publicado: (2025)
por: Christoph, Daniel, et al.
Publicado: (2025)
SpecRover: Code Intent Extraction via LLMs
por: Ruan, Haifeng, et al.
Publicado: (2024)
por: Ruan, Haifeng, et al.
Publicado: (2024)
CodeWatcher: IDE Telemetry Data Extraction Tool for Understanding Coding Interactions with LLMs
por: Basha, Manaal, et al.
Publicado: (2025)
por: Basha, Manaal, et al.
Publicado: (2025)
Transformer-Based Extraction of Statutory Definitions from the U.S. Code
por: Hosabettu, Arpana, et al.
Publicado: (2025)
por: Hosabettu, Arpana, et al.
Publicado: (2025)
Seemingly Simple Planning Problems are Computationally Challenging: The Countdown Game
por: Katz, Michael, et al.
Publicado: (2025)
por: Katz, Michael, et al.
Publicado: (2025)
KnowCoder: Coding Structured Knowledge into LLMs for Universal Information Extraction
por: Li, Zixuan, et al.
Publicado: (2024)
por: Li, Zixuan, et al.
Publicado: (2024)
Ten Challenging Problems in Federated Foundation Models
por: Fan, Tao, et al.
Publicado: (2025)
por: Fan, Tao, et al.
Publicado: (2025)
RoboWits: Unexpected Challenges for Robotic Creative Problem Solving
por: Lin, Chunru, et al.
Publicado: (2026)
por: Lin, Chunru, et al.
Publicado: (2026)
LLM Multi-Agent Systems: Challenges and Open Problems
por: Han, Shanshan, et al.
Publicado: (2024)
por: Han, Shanshan, et al.
Publicado: (2024)
Culturally Responsive Artificial Intelligence -- Problems, Challenges and Solutions
por: Ożegalska-Łukasik, Natalia, et al.
Publicado: (2023)
por: Ożegalska-Łukasik, Natalia, et al.
Publicado: (2023)
Code-Based English Models Surprising Performance on Chinese QA Pair Extraction Task
por: Zheng, Linghan, et al.
Publicado: (2024)
por: Zheng, Linghan, et al.
Publicado: (2024)
Beyond Embeddings: Interpretable Feature Extraction for Binary Code Similarity
por: Gagnon, Charles E., et al.
Publicado: (2025)
por: Gagnon, Charles E., et al.
Publicado: (2025)
Multi-Agent Reinforcement Learning for Energy Networks: Computational Challenges, Progress and Open Problems
por: Keren, Sarah, et al.
Publicado: (2024)
por: Keren, Sarah, et al.
Publicado: (2024)
Problem Solved? Information Extraction Design Space for Layout-Rich Documents using LLMs
por: Colakoglu, Gaye, et al.
Publicado: (2025)
por: Colakoglu, Gaye, et al.
Publicado: (2025)
Maximizing Relation Extraction Potential: A Data-Centric Study to Unveil Challenges and Opportunities
por: Swarup, Anushka, et al.
Publicado: (2024)
por: Swarup, Anushka, et al.
Publicado: (2024)
Queueing, Predictions, and LLMs: Challenges and Open Problems
por: Mitzenmacher, Michael, et al.
Publicado: (2025)
por: Mitzenmacher, Michael, et al.
Publicado: (2025)
Agent RL Scaling Law: Agent RL with Spontaneous Code Execution for Mathematical Problem Solving
por: Mai, Xinji, et al.
Publicado: (2025)
por: Mai, Xinji, et al.
Publicado: (2025)
Evaluating Code Generation of LLMs in Advanced Computer Science Problems
por: Catir, Emir, et al.
Publicado: (2025)
por: Catir, Emir, et al.
Publicado: (2025)
Hallucination by Code Generation LLMs: Taxonomy, Benchmarks, Mitigation, and Challenges
por: Lee, Yunseo, et al.
Publicado: (2025)
por: Lee, Yunseo, et al.
Publicado: (2025)
KodCode: A Diverse, Challenging, and Verifiable Synthetic Dataset for Coding
por: Xu, Zhangchen, et al.
Publicado: (2025)
por: Xu, Zhangchen, et al.
Publicado: (2025)
HARDMath: A Benchmark Dataset for Challenging Problems in Applied Mathematics
por: Fan, Jingxuan, et al.
Publicado: (2024)
por: Fan, Jingxuan, et al.
Publicado: (2024)
Optimizing Sensor Redundancy in Sequential Decision-Making Problems
por: Nüßlein, Jonas, et al.
Publicado: (2024)
por: Nüßlein, Jonas, et al.
Publicado: (2024)
KnowCoder-X: Boosting Multilingual Information Extraction via Code
por: Zuo, Yuxin, et al.
Publicado: (2024)
por: Zuo, Yuxin, et al.
Publicado: (2024)
Keyword Extraction, and Aspect Classification in Sinhala, English, and Code-Mixed Content
por: Rizvi, F. A., et al.
Publicado: (2025)
por: Rizvi, F. A., et al.
Publicado: (2025)
ImpReSS: Implicit Recommender System for Support Conversations
por: Haller, Omri, et al.
Publicado: (2025)
por: Haller, Omri, et al.
Publicado: (2025)
No More Blind Spots: Learning Vision-Based Omnidirectional Bipedal Locomotion for Challenging Terrain
por: Gadde, Mohitvishnu S., et al.
Publicado: (2025)
por: Gadde, Mohitvishnu S., et al.
Publicado: (2025)
Ejemplares similares
-
Sample-Efficient Language Modeling with Linear Attention and Lightweight Enhancements
por: Haller, Patrick, et al.
Publicado: (2025) -
What Matters in Linearizing Language Models? A Comparative Study of Architecture, Scale, and Task Adaptation
por: Haller, Patrick, et al.
Publicado: (2025) -
Fabricator: An Open Source Toolkit for Generating Labeled Training Data with Teacher LLMs
por: Golde, Jonas, et al.
Publicado: (2023) -
PISA-Bench: The PISA Index as a Multilingual and Multimodal Metric for the Evaluation of Vision-Language Models
por: Haller, Patrick, et al.
Publicado: (2025) -
BabyHGRN: Exploring RNNs for Sample-Efficient Training of Language Models
por: Haller, Patrick, et al.
Publicado: (2024)