SAGE:Specification-Aware Grammar Extraction for Automated Test Case Generation with LLMs
Fuente:
arXiv
Guardado en:
| Autores principales: | Aditi, Park, Hyunwoo, Sung, Sicheol, Han, Yo-Sub, Ko, Sang-Ki |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
LogiCase: Effective Test Case Generation from Logical Description in Competitive Programming
por: Sung, Sicheol, et al.
Publicado: (2025)
por: Sung, Sicheol, et al.
Publicado: (2025)
A Framework for Quantum Finite-State Languages with Density Mapping
por: Baik, SeungYeop, et al.
Publicado: (2024)
por: Baik, SeungYeop, et al.
Publicado: (2024)
Repairing Regex Vulnerabilities via Localization-Guided Instructions
por: Sung, Sicheol, et al.
Publicado: (2025)
por: Sung, Sicheol, et al.
Publicado: (2025)
From Intuition to Calibrated Judgment: A Rubric-Based Expert-Panel Study of Human Detection of LLM-Generated Korean Text
por: Park, Shinwoo, et al.
Publicado: (2026)
por: Park, Shinwoo, et al.
Publicado: (2026)
Linguistics-Aware Non-Distortionary LLM Watermarking
por: Park, Shinwoo, et al.
Publicado: (2026)
por: Park, Shinwoo, et al.
Publicado: (2026)
A Linguistics-Aware LLM Watermarking via Syntactic Predictability
por: Park, Shinwoo, et al.
Publicado: (2025)
por: Park, Shinwoo, et al.
Publicado: (2025)
TRAPDOC: Deceiving LLM Users by Injecting Imperceptible Phantom Tokens into Documents
por: Jin, Hyundong, et al.
Publicado: (2025)
por: Jin, Hyundong, et al.
Publicado: (2025)
ContractEval: A Benchmark for Evaluating Contract-Satisfying Assertions in Code Generation
por: Lim, Soohan, et al.
Publicado: (2025)
por: Lim, Soohan, et al.
Publicado: (2025)
DLM-SWAI: Steering Diffusion Language Models Before They Unmask
por: An, Hyeseon, et al.
Publicado: (2026)
por: An, Hyeseon, et al.
Publicado: (2026)
ReSyn: A Generalized Recursive Regular Expression Synthesis Framework
por: Kim, Seongmin, et al.
Publicado: (2026)
por: Kim, Seongmin, et al.
Publicado: (2026)
Adaptive Steering and Remasking for Safe Generation in Diffusion Language Models
por: Lee, Yejin, et al.
Publicado: (2026)
por: Lee, Yejin, et al.
Publicado: (2026)
EPIC: Efficient and Parallel Inference under CFG Constraints for Diffusion Language Models
por: Jin, Hyundong, et al.
Publicado: (2026)
por: Jin, Hyundong, et al.
Publicado: (2026)
NCO: A Versatile Plug-in for Handling Negative Constraints in Decoding
por: Jin, Hyundong, et al.
Publicado: (2026)
por: Jin, Hyundong, et al.
Publicado: (2026)
CodeComplex: Dataset for Worst-Case Time Complexity Prediction
por: Baik, Seung-Yeop, et al.
Publicado: (2024)
por: Baik, Seung-Yeop, et al.
Publicado: (2024)
Steering Language Models Before They Speak: Logit-Level Interventions
por: An, Hyeseon, et al.
Publicado: (2026)
por: An, Hyeseon, et al.
Publicado: (2026)
KatFishNet: Detecting LLM-Generated Korean Text through Linguistic Feature Analysis
por: Park, Shinwoo, et al.
Publicado: (2025)
por: Park, Shinwoo, et al.
Publicado: (2025)
SAGE: Shaping Anchors for Guided Exploration in RLVR of LLMs
por: Lee, Chanuk, et al.
Publicado: (2026)
por: Lee, Chanuk, et al.
Publicado: (2026)
ECO: Enhanced Code Optimization via Performance-Aware Prompting for Code-LLMs
por: Kim, Su-Hyeon, et al.
Publicado: (2025)
por: Kim, Su-Hyeon, et al.
Publicado: (2025)
Can LLMs Help Create Grammar?: Automating Grammar Creation for Endangered Languages with In-Context Learning
por: Spencer, Piyapath T, et al.
Publicado: (2024)
por: Spencer, Piyapath T, et al.
Publicado: (2024)
Controlling Language Confusion in Multilingual LLMs
por: Lee, Nahyun, et al.
Publicado: (2025)
por: Lee, Nahyun, et al.
Publicado: (2025)
ImpossibleBench: Measuring LLMs' Propensity of Exploiting Test Cases
por: Zhong, Ziqian, et al.
Publicado: (2025)
por: Zhong, Ziqian, et al.
Publicado: (2025)
RV-HATE: Reinforced Multi-Module Voting for Implicit Hate Speech Detection
por: Lee, Yejin, et al.
Publicado: (2025)
por: Lee, Yejin, et al.
Publicado: (2025)
Linguistic Generalizability of Test-Time Scaling in Mathematical Reasoning
por: Son, Guijin, et al.
Publicado: (2025)
por: Son, Guijin, et al.
Publicado: (2025)
Towards Prompt Generalization: Grammar-aware Cross-Prompt Automated Essay Scoring
por: Do, Heejin, et al.
Publicado: (2025)
por: Do, Heejin, et al.
Publicado: (2025)
Domain-Specific Shorthand for Generation Based on Context-Free Grammar
por: Kanyuka, Andriy, et al.
Publicado: (2024)
por: Kanyuka, Andriy, et al.
Publicado: (2024)
Grammars of Formal Uncertainty: When to Trust LLMs in Automated Reasoning Tasks
por: Ganguly, Debargha, et al.
Publicado: (2025)
por: Ganguly, Debargha, et al.
Publicado: (2025)
AmpleHate: Amplifying the Attention for Versatile Implicit Hate Detection
por: Lee, Yejin, et al.
Publicado: (2025)
por: Lee, Yejin, et al.
Publicado: (2025)
Accurate and Efficient Statistical Testing for Word Semantic Breadth
por: Ehara, Yo
Publicado: (2026)
por: Ehara, Yo
Publicado: (2026)
KAIO: A Collection of More Challenging Korean Questions
por: Lee, Nahyun, et al.
Publicado: (2025)
por: Lee, Nahyun, et al.
Publicado: (2025)
SAGE: Scalable Automated Robustness Augmentation for LLM Knowledge Evaluation
por: Li, Xiaoyuan, et al.
Publicado: (2026)
por: Li, Xiaoyuan, et al.
Publicado: (2026)
General LLMs as Instructors for Domain-Specific LLMs: A Sequential Fusion Method to Integrate Extraction and Editing
por: Zhang, Xin, et al.
Publicado: (2024)
por: Zhang, Xin, et al.
Publicado: (2024)
SAGE: Steering Dialog Generation with Future-Aware State-Action Augmentation
por: Zhang, Yizhe, et al.
Publicado: (2025)
por: Zhang, Yizhe, et al.
Publicado: (2025)
Multi-Step Reasoning in Korean and the Emergent Mirage
por: Son, Guijin, et al.
Publicado: (2025)
por: Son, Guijin, et al.
Publicado: (2025)
Understand, Solve and Translate: Bridging the Multilingual Mathematical Reasoning Gap
por: Ko, Hyunwoo, et al.
Publicado: (2025)
por: Ko, Hyunwoo, et al.
Publicado: (2025)
Harnessing Generative LLMs for Enhanced Financial Event Entity Extraction Performance
por: Choi, Soo-joon, et al.
Publicado: (2025)
por: Choi, Soo-joon, et al.
Publicado: (2025)
Grammaticality Judgments in Humans and Language Models: Revisiting Generative Grammar with LLMs
por: Johnsen, Lars G. B.
Publicado: (2025)
por: Johnsen, Lars G. B.
Publicado: (2025)
Polishing Every Facet of the GEM: Testing Linguistic Competence of LLMs and Humans in Korean
por: Kim, SungHo, et al.
Publicado: (2025)
por: Kim, SungHo, et al.
Publicado: (2025)
Grammar and Gameplay-aligned RL for Game Description Generation with LLMs
por: Tanaka, Tsunehiko, et al.
Publicado: (2025)
por: Tanaka, Tsunehiko, et al.
Publicado: (2025)
More Aligned, Less Diverse? Analyzing the Grammar and Lexicon of Two Generations of LLMs
por: Gude, Adrián, et al.
Publicado: (2026)
por: Gude, Adrián, et al.
Publicado: (2026)
RegexPSPACE: A Benchmark for Evaluating LLM Reasoning on PSPACE-complete Regex Problems
por: Jin, Hyundong, et al.
Publicado: (2025)
por: Jin, Hyundong, et al.
Publicado: (2025)
Ejemplares similares
-
LogiCase: Effective Test Case Generation from Logical Description in Competitive Programming
por: Sung, Sicheol, et al.
Publicado: (2025) -
A Framework for Quantum Finite-State Languages with Density Mapping
por: Baik, SeungYeop, et al.
Publicado: (2024) -
Repairing Regex Vulnerabilities via Localization-Guided Instructions
por: Sung, Sicheol, et al.
Publicado: (2025) -
From Intuition to Calibrated Judgment: A Rubric-Based Expert-Panel Study of Human Detection of LLM-Generated Korean Text
por: Park, Shinwoo, et al.
Publicado: (2026) -
Linguistics-Aware Non-Distortionary LLM Watermarking
por: Park, Shinwoo, et al.
Publicado: (2026)