Position Paper: Generalized grammar rules and structure-based generalization beyond classical equivariance for lexical tasks and transduction
Fuente:
arXiv
Guardado en:
| Autores principales: | Petrache, Mircea, Trivedi, Shubhendu |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Approximation-Generalization Trade-offs under (Approximate) Group Equivariance
por: Petrache, Mircea, et al.
Publicado: (2023)
por: Petrache, Mircea, et al.
Publicado: (2023)
Recurrent Equivariant Constraint Modulation: Learning Per-Layer Symmetry Relaxation from Data
por: Pertigkiozoglou, Stefanos, et al.
Publicado: (2026)
por: Pertigkiozoglou, Stefanos, et al.
Publicado: (2026)
Symmetry-Based Structured Matrices for Efficient Approximately Equivariant Networks
por: Samudre, Ashwin, et al.
Publicado: (2024)
por: Samudre, Ashwin, et al.
Publicado: (2024)
Generating with Confidence: Uncertainty Quantification for Black-box Large Language Models
por: Lin, Zhen, et al.
Publicado: (2023)
por: Lin, Zhen, et al.
Publicado: (2023)
On Universality of Deep Equivariant Networks
por: Pacini, Marco, et al.
Publicado: (2025)
por: Pacini, Marco, et al.
Publicado: (2025)
Watermarking Degrades Alignment in Language Models: Analysis and Mitigation
por: Verma, Apurv, et al.
Publicado: (2025)
por: Verma, Apurv, et al.
Publicado: (2025)
Transduce: learning transduction grammars for string transformation
por: Frydman, Francis, et al.
Publicado: (2023)
por: Frydman, Francis, et al.
Publicado: (2023)
Contextualized Sequence Likelihood: Enhanced Confidence Scores for Natural Language Generation
por: Lin, Zhen, et al.
Publicado: (2024)
por: Lin, Zhen, et al.
Publicado: (2024)
On the Need to Align Intent and Implementation in Uncertainty Quantification for Machine Learning
por: Trivedi, Shubhendu, et al.
Publicado: (2025)
por: Trivedi, Shubhendu, et al.
Publicado: (2025)
Weak-to-Strong Generalization beyond Accuracy: a Pilot Study in Safety, Toxicity, and Legal Reasoning
por: Ye, Ruimeng, et al.
Publicado: (2024)
por: Ye, Ruimeng, et al.
Publicado: (2024)
COLD-Steer: Steering Large Language Models via In-Context One-step Learning Dynamics
por: Sharma, Kartik, et al.
Publicado: (2026)
por: Sharma, Kartik, et al.
Publicado: (2026)
Efficient Knowledge Probing of Large Language Models by Adapting Pre-trained Embeddings
por: Sharma, Kartik, et al.
Publicado: (2025)
por: Sharma, Kartik, et al.
Publicado: (2025)
Insertion Language Models: Sequence Generation with Arbitrary-Position Insertions
por: Patel, Dhruvesh, et al.
Publicado: (2025)
por: Patel, Dhruvesh, et al.
Publicado: (2025)
DeepReview: Improving LLM-based Paper Review with Human-like Deep Thinking Process
por: Zhu, Minjun, et al.
Publicado: (2025)
por: Zhu, Minjun, et al.
Publicado: (2025)
Language models scale reliably with over-training and on downstream tasks
por: Gadre, Samir Yitzhak, et al.
Publicado: (2024)
por: Gadre, Samir Yitzhak, et al.
Publicado: (2024)
Deep Prompt Multi-task Network for Abuse Language Detection
por: Zhu, Jian, et al.
Publicado: (2024)
por: Zhu, Jian, et al.
Publicado: (2024)
Transformers need glasses! Information over-squashing in language tasks
por: Barbero, Federico, et al.
Publicado: (2024)
por: Barbero, Federico, et al.
Publicado: (2024)
Harnessing Test-time Adaptation for NLU tasks Involving Dialects of English
por: Nguyen, Duke, et al.
Publicado: (2025)
por: Nguyen, Duke, et al.
Publicado: (2025)
Learning beyond Teacher: Generalized On-Policy Distillation with Reward Extrapolation
por: Yang, Wenkai, et al.
Publicado: (2026)
por: Yang, Wenkai, et al.
Publicado: (2026)
Tackling prediction tasks in relational databases with LLMs
por: Wydmuch, Marek, et al.
Publicado: (2024)
por: Wydmuch, Marek, et al.
Publicado: (2024)
RuAG: Learned-rule-augmented Generation for Large Language Models
por: Zhang, Yudi, et al.
Publicado: (2024)
por: Zhang, Yudi, et al.
Publicado: (2024)
RoLegalGEC: Legal Domain Grammatical Error Detection and Correction Dataset for Romanian
por: Timpuriu, Mircea, et al.
Publicado: (2026)
por: Timpuriu, Mircea, et al.
Publicado: (2026)
A Compressive-Expressive Communication Framework for Compositional Representations
por: Elberg, Rafael, et al.
Publicado: (2025)
por: Elberg, Rafael, et al.
Publicado: (2025)
nach0-pc: Multi-task Language Model with Molecular Point Cloud Encoder
por: Kuznetsov, Maksim, et al.
Publicado: (2024)
por: Kuznetsov, Maksim, et al.
Publicado: (2024)
Scaling behavior of large language models in emotional safety classification across sizes and tasks
por: Pinzuti, Edoardo, et al.
Publicado: (2025)
por: Pinzuti, Edoardo, et al.
Publicado: (2025)
Do different prompting methods yield a common task representation in language models?
por: Davidson, Guy, et al.
Publicado: (2025)
por: Davidson, Guy, et al.
Publicado: (2025)
Unlearning as multi-task optimization: A normalized gradient difference approach with an adaptive learning rate
por: Bu, Zhiqi, et al.
Publicado: (2024)
por: Bu, Zhiqi, et al.
Publicado: (2024)
Which Side Are You On? A Multi-task Dataset for End-to-End Argument Summarisation and Evaluation
por: Li, Hao, et al.
Publicado: (2024)
por: Li, Hao, et al.
Publicado: (2024)
Beyond instruction-conditioning, MoTE: Mixture of Task Experts for Multi-task Embedding Models
por: Romero, Miguel, et al.
Publicado: (2025)
por: Romero, Miguel, et al.
Publicado: (2025)
Adaptively profiling models with task elicitation
por: Brown, Davis, et al.
Publicado: (2025)
por: Brown, Davis, et al.
Publicado: (2025)
Selective Rotary Position Embedding
por: Movahedi, Sajad, et al.
Publicado: (2025)
por: Movahedi, Sajad, et al.
Publicado: (2025)
On the Geometry of Positional Encodings in Transformers
por: Cirrincione, Giansalvo
Publicado: (2026)
por: Cirrincione, Giansalvo
Publicado: (2026)
Position Information Emerges in Causal Transformers Without Positional Encodings via Similarity of Nearby Embeddings
por: Zuo, Chunsheng, et al.
Publicado: (2024)
por: Zuo, Chunsheng, et al.
Publicado: (2024)
From Small to Large Language Models: Revisiting the Federalist Papers
por: Jeong, So Won, et al.
Publicado: (2025)
por: Jeong, So Won, et al.
Publicado: (2025)
Inner Speech as Behavior Guides: Steerable Imitation of Diverse Behaviors for Human-AI coordination
por: Trivedi, Rakshit, et al.
Publicado: (2026)
por: Trivedi, Rakshit, et al.
Publicado: (2026)
Impact of automatic speech recognition quality on Alzheimer's disease detection from spontaneous speech: a reproducible benchmark study with lexical modeling and statistical validation
por: Samanta, Himadri S
Publicado: (2026)
por: Samanta, Himadri S
Publicado: (2026)
Evaluating the fairness of task-adaptive pretraining on unlabeled test data before few-shot text classification
por: Dubey, Kush
Publicado: (2024)
por: Dubey, Kush
Publicado: (2024)
RubricEM: Meta-RL with Rubric-guided Policy Decomposition beyond Verifiable Rewards
por: Li, Gaotang, et al.
Publicado: (2026)
por: Li, Gaotang, et al.
Publicado: (2026)
Looking beyond the next token
por: Thankaraj, Abitha, et al.
Publicado: (2025)
por: Thankaraj, Abitha, et al.
Publicado: (2025)
Bifocal Attention: Harmonizing Geometric and Spectral Positional Embeddings for Algorithmic Generalization
por: Awadhiya, Kanishk
Publicado: (2026)
por: Awadhiya, Kanishk
Publicado: (2026)
Ejemplares similares
-
Approximation-Generalization Trade-offs under (Approximate) Group Equivariance
por: Petrache, Mircea, et al.
Publicado: (2023) -
Recurrent Equivariant Constraint Modulation: Learning Per-Layer Symmetry Relaxation from Data
por: Pertigkiozoglou, Stefanos, et al.
Publicado: (2026) -
Symmetry-Based Structured Matrices for Efficient Approximately Equivariant Networks
por: Samudre, Ashwin, et al.
Publicado: (2024) -
Generating with Confidence: Uncertainty Quantification for Black-box Large Language Models
por: Lin, Zhen, et al.
Publicado: (2023) -
On Universality of Deep Equivariant Networks
por: Pacini, Marco, et al.
Publicado: (2025)