Merge-Bench: Resolve Merge Conflicts with Large Language Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Schesch, Benedikt, Ernst, Michael D. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Automated Bug Triaging using Instruction-Tuned Large Language Models
por: Kiashemshaki, Kiana, et al.
Publicado: (2025)
por: Kiashemshaki, Kiana, et al.
Publicado: (2025)
Entropy-Based Measurement of Value Drift and Alignment Work in Large Language Models
por: Fadli, Samih
Publicado: (2025)
por: Fadli, Samih
Publicado: (2025)
Mixup Model Merge: Enhancing Model Merging Performance through Randomized Linear Interpolation
por: Zhou, Yue, et al.
Publicado: (2025)
por: Zhou, Yue, et al.
Publicado: (2025)
Evaluating the Systematic Reasoning Abilities of Large Language Models through Graph Coloring
por: Heyman, Alex, et al.
Publicado: (2025)
por: Heyman, Alex, et al.
Publicado: (2025)
Transformer Scalability Crisis: The First Comprehensive Empirical Analysis of Performance Walls in Modern Language Models
por: Moghadasi, Mahdi Naser, et al.
Publicado: (2026)
por: Moghadasi, Mahdi Naser, et al.
Publicado: (2026)
Learned Relay Representations for Forward-Thinking Discrete Diffusion Models
por: Rozonoyer, Benjamin, et al.
Publicado: (2026)
por: Rozonoyer, Benjamin, et al.
Publicado: (2026)
Variance Is Not Importance: Structural Analysis of Transformer Compressibility Across Model Scales
por: Salfati, Samuel
Publicado: (2026)
por: Salfati, Samuel
Publicado: (2026)
QuAnTS: Question Answering on Time Series
por: Divo, Felix, et al.
Publicado: (2025)
por: Divo, Felix, et al.
Publicado: (2025)
Latent Instruction Representation Alignment: defending against jailbreaks, backdoors and undesired knowledge in LLMs
por: Easley, Eric, et al.
Publicado: (2026)
por: Easley, Eric, et al.
Publicado: (2026)
Descriptive Collision in Sparse Autoencoder Auto-Interpretability: When One Explanation Describes Many Features
por: McCann, Jordan F.
Publicado: (2026)
por: McCann, Jordan F.
Publicado: (2026)
Super Apriel: One Checkpoint, Many Speeds
por: Labs, SLAM, et al.
Publicado: (2026)
por: Labs, SLAM, et al.
Publicado: (2026)
Why LoRA Resists Label Noise: A Theoretical Framework for Noise-Robust Parameter-Efficient Fine-Tuning
por: Steele, Brady
Publicado: (2026)
por: Steele, Brady
Publicado: (2026)
TensorLens: End-to-End Transformer Analysis via High-Order Attention Tensors
por: Atad, Ido Andrew, et al.
Publicado: (2026)
por: Atad, Ido Andrew, et al.
Publicado: (2026)
ACE: Exploring Activation Cosine Similarity and Variance for Accurate and Calibration-Efficient LLM Pruning
por: Mi, Zhendong, et al.
Publicado: (2025)
por: Mi, Zhendong, et al.
Publicado: (2025)
Scalable GPU-Accelerated Euler Characteristic Curves: Optimization and Differentiable Learning for PyTorch
por: Saxena, Udit
Publicado: (2025)
por: Saxena, Udit
Publicado: (2025)
Synergy over Discrepancy: A Partition-Based Approach to Multi-Domain LLM Fine-Tuning
por: Ye, Hua, et al.
Publicado: (2025)
por: Ye, Hua, et al.
Publicado: (2025)
KerZOO: Kernel Function Informed Zeroth-Order Optimization for Accurate and Accelerated LLM Fine-Tuning
por: Mi, Zhendong, et al.
Publicado: (2025)
por: Mi, Zhendong, et al.
Publicado: (2025)
Revisiting LRP: Positional Attribution as the Missing Ingredient for Transformer Explainability
por: Bakish, Yarden, et al.
Publicado: (2025)
por: Bakish, Yarden, et al.
Publicado: (2025)
Memory Bank Compression for Continual Adaptation of Large Language Models
por: Katraouras, Thomas, et al.
Publicado: (2026)
por: Katraouras, Thomas, et al.
Publicado: (2026)
The Geometry of Thought: How Scale Restructures Reasoning In Large Language Models
por: Anderson, Samuel Cyrenius
Publicado: (2026)
por: Anderson, Samuel Cyrenius
Publicado: (2026)
PoTS: Proof-of-Training-Steps for Backdoor Detection in Large Language Models
por: Seddik, Issam, et al.
Publicado: (2025)
por: Seddik, Issam, et al.
Publicado: (2025)
Language Models Are Implicitly Continuous
por: Marro, Samuele, et al.
Publicado: (2025)
por: Marro, Samuele, et al.
Publicado: (2025)
Reasoning Large Language Model Errors Arise from Hallucinating Critical Problem Features
por: Heyman, Alex, et al.
Publicado: (2025)
por: Heyman, Alex, et al.
Publicado: (2025)
The Anti-Ouroboros Effect: Emergent Resilience in Large Language Models from Recursive Selective Feedback
por: Adapala, Sai Teja Reddy
Publicado: (2025)
por: Adapala, Sai Teja Reddy
Publicado: (2025)
Rethinking Addressing in Language Models via Contexualized Equivariant Positional Encoding
por: Zhu, Jiajun, et al.
Publicado: (2025)
por: Zhu, Jiajun, et al.
Publicado: (2025)
Latent Cache Flow: Model-to-Model Communication Without Text
por: Rossi, Maximillian, et al.
Publicado: (2026)
por: Rossi, Maximillian, et al.
Publicado: (2026)
Combining Language and Topic Models for Hierarchical Text Classification
por: Toit, Jaco du, et al.
Publicado: (2025)
por: Toit, Jaco du, et al.
Publicado: (2025)
Recurrent Memory-Augmented Transformers with Chunked Attention for Long-Context Language Modeling
por: Kashyap, Ankit
Publicado: (2025)
por: Kashyap, Ankit
Publicado: (2025)
From Syntax to Semantics: Unveiling the Emergence of Chirality in SMILES Translation Models
por: Li, Zehao, et al.
Publicado: (2026)
por: Li, Zehao, et al.
Publicado: (2026)
Relating Misfit to Gain in Weak-to-Strong Generalization Beyond the Squared Loss
por: Mulgund, Abhijeet, et al.
Publicado: (2025)
por: Mulgund, Abhijeet, et al.
Publicado: (2025)
Kronecker Embeddings: Byte-Level Structured Token Representations for Parameter-Efficient Language Models
por: Shravan, Rohan
Publicado: (2026)
por: Shravan, Rohan
Publicado: (2026)
How Language Models Process Out-of-Distribution Inputs: A Two-Pathway Framework
por: Saghir, Hamidreza
Publicado: (2026)
por: Saghir, Hamidreza
Publicado: (2026)
Neural Activation Patterns Across Language Model Architectures: A Comprehensive Analysis of Cognitive Task Performance
por: Naser-Moghadasi, Mahdi, et al.
Publicado: (2026)
por: Naser-Moghadasi, Mahdi, et al.
Publicado: (2026)
Future Token Prediction -- Causal Language Modelling with Per-Token Semantic State Vector for Multi-Token Prediction
por: Walker, Nicholas
Publicado: (2024)
por: Walker, Nicholas
Publicado: (2024)
Your Pretrained Model Tells the Difficulty Itself: A Self-Adaptive Curriculum Learning Paradigm for Natural Language Understanding
por: Feng, Qi, et al.
Publicado: (2025)
por: Feng, Qi, et al.
Publicado: (2025)
Sarcasm Detection in a Less-Resourced Language
por: Đoković, Lazar, et al.
Publicado: (2024)
por: Đoković, Lazar, et al.
Publicado: (2024)
ConfProBench: A Confidence Evaluation Benchmark for MLLM-Based Process Judges
por: Zhou, Yue, et al.
Publicado: (2025)
por: Zhou, Yue, et al.
Publicado: (2025)
Advancing Transformer Architecture in Long-Context Large Language Models: A Comprehensive Survey
por: Huang, Yunpeng, et al.
Publicado: (2023)
por: Huang, Yunpeng, et al.
Publicado: (2023)
Improving ML Training Data with Gold-Standard Quality Metrics
por: Barrett, Leslie, et al.
Publicado: (2025)
por: Barrett, Leslie, et al.
Publicado: (2025)
Towards Alignment-Centric Paradigm: A Survey of Instruction Tuning in Large Language Models
por: Han, Xudong, et al.
Publicado: (2025)
por: Han, Xudong, et al.
Publicado: (2025)
Ejemplares similares
-
Automated Bug Triaging using Instruction-Tuned Large Language Models
por: Kiashemshaki, Kiana, et al.
Publicado: (2025) -
Entropy-Based Measurement of Value Drift and Alignment Work in Large Language Models
por: Fadli, Samih
Publicado: (2025) -
Mixup Model Merge: Enhancing Model Merging Performance through Randomized Linear Interpolation
por: Zhou, Yue, et al.
Publicado: (2025) -
Evaluating the Systematic Reasoning Abilities of Large Language Models through Graph Coloring
por: Heyman, Alex, et al.
Publicado: (2025) -
Transformer Scalability Crisis: The First Comprehensive Empirical Analysis of Performance Walls in Modern Language Models
por: Moghadasi, Mahdi Naser, et al.
Publicado: (2026)