Saved in:
| Main Authors: | Bradley, Tai-Danae, Vigneaux, Juan Pablo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2501.06662 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Mathematical Theory of Discursive Networks
by: Gutiérrez, Juan B.
Published: (2025)
by: Gutiérrez, Juan B.
Published: (2025)
On internal categories and crossed objects in the category of monoids
by: Pirashvili, Ilia
Published: (2024)
by: Pirashvili, Ilia
Published: (2024)
Recipient Profiling: Predicting Characteristics from Messages
by: Borquez, Martin, et al.
Published: (2024)
by: Borquez, Martin, et al.
Published: (2024)
Data and AI governance: Promoting equity, ethics, and fairness in large language models
by: Abhishek, Alok, et al.
Published: (2025)
by: Abhishek, Alok, et al.
Published: (2025)
The 99% Success Paradox: When Near-Perfect Retrieval Equals Random Selection
by: Repantis, Vyzantinos, et al.
Published: (2026)
by: Repantis, Vyzantinos, et al.
Published: (2026)
The Lossy Horizon: Error-Bounded Predictive Coding for Lossy Text Compression (Episode I)
by: Aghanya, Nnamdi, et al.
Published: (2025)
by: Aghanya, Nnamdi, et al.
Published: (2025)
Asymptotic Semantic Collapse in Hierarchical Optimization
by: Alpay, Faruk, et al.
Published: (2026)
by: Alpay, Faruk, et al.
Published: (2026)
Measuring text summarization factuality using atomic facts entailment metrics in the context of retrieval augmented generation
by: Kriman, N. E.
Published: (2024)
by: Kriman, N. E.
Published: (2024)
UM_FHS at TREC 2024 PLABA: Exploration of Fine-tuning and AI agent approach for plain language adaptations of biomedical text
by: Kocbek, Primoz, et al.
Published: (2025)
by: Kocbek, Primoz, et al.
Published: (2025)
PLUGH: A Benchmark for Spatial Understanding and Reasoning in Large Language Models
by: Tikhonov, Alexey
Published: (2024)
by: Tikhonov, Alexey
Published: (2024)
Nerves of enriched categories via necklaces
by: Mertens, Arne
Published: (2024)
by: Mertens, Arne
Published: (2024)
Unveiling factors influencing judgment variation in Sentiment Analysis with Natural Language Processing and Statistics
by: Kellert, Olga, et al.
Published: (2024)
by: Kellert, Olga, et al.
Published: (2024)
An Iterative Optimizing Framework for Radiology Report Summarization with ChatGPT
by: Ma, Chong, et al.
Published: (2023)
by: Ma, Chong, et al.
Published: (2023)
Large Language Models Report Subjective Experience Under Self-Referential Processing
by: Berg, Cameron, et al.
Published: (2025)
by: Berg, Cameron, et al.
Published: (2025)
From MTEB to MTOB: Retrieval-Augmented Classification for Descriptive Grammars
by: Kornilov, Albert, et al.
Published: (2024)
by: Kornilov, Albert, et al.
Published: (2024)
Entropies associated with orbits of finite groups
by: Leal, Ryan, et al.
Published: (2025)
by: Leal, Ryan, et al.
Published: (2025)
Chatbots put to the test in math and logic problems: A preliminary comparison and assessment of ChatGPT-3.5, ChatGPT-4, and Google Bard
by: Plevris, Vagelis, et al.
Published: (2023)
by: Plevris, Vagelis, et al.
Published: (2023)
Prompt Tuned Embedding Classification for Multi-Label Industry Sector Allocation
by: Buchner, Valentin Leonhard, et al.
Published: (2023)
by: Buchner, Valentin Leonhard, et al.
Published: (2023)
Reference-Guided Verdict: LLMs-as-Judges in Automatic Evaluation of Free-Form QA
by: Badshah, Sher, et al.
Published: (2024)
by: Badshah, Sher, et al.
Published: (2024)
SERC: LDPC-Inspired Semantic Error Correction for Retrieval-Augmented Generation
by: Kim, Gyumin, et al.
Published: (2026)
by: Kim, Gyumin, et al.
Published: (2026)
Technical Report on the Pangram AI-Generated Text Classifier
by: Emi, Bradley, et al.
Published: (2024)
by: Emi, Bradley, et al.
Published: (2024)
Co-NAML-LSTUR: A Combined Model with Attentive Multi-View Learning and Long- and Short-term User Representations for News Recommendation
by: Nguyen, Minh Hoang, et al.
Published: (2025)
by: Nguyen, Minh Hoang, et al.
Published: (2025)
PaperAudit-Bench: Benchmarking Error Detection in Research Papers for Critical Automated Peer Review
by: Tu, Songjun, et al.
Published: (2026)
by: Tu, Songjun, et al.
Published: (2026)
SocialX: A Modular Platform for Multi-Source Big Data Research in Indonesia
by: Saputra, Muhammad Apriandito Arya, et al.
Published: (2026)
by: Saputra, Muhammad Apriandito Arya, et al.
Published: (2026)
Tethered Reasoning: Decoupling Entropy from Hallucination in Quantized LLMs via Manifold Steering
by: Atkinson, Craig
Published: (2026)
by: Atkinson, Craig
Published: (2026)
LLM-Viterbi: Semantic-Aware Decoding for Convolutional Codes
by: Li, Zhengtong, et al.
Published: (2026)
by: Li, Zhengtong, et al.
Published: (2026)
Inference acceleration for large language models using "stairs" assisted greedy generation
by: Grigaliūnas, Domas, et al.
Published: (2024)
by: Grigaliūnas, Domas, et al.
Published: (2024)
When Many-Shot Prompting Fails: An Empirical Study of LLM Code Translation
by: Oskooei, Amirkia Rafiei, et al.
Published: (2025)
by: Oskooei, Amirkia Rafiei, et al.
Published: (2025)
SHARP: Social Harm Analysis via Risk Profiles for Measuring Inequities in Large Language Models
by: Abhishek, Alok, et al.
Published: (2026)
by: Abhishek, Alok, et al.
Published: (2026)
BEATS: Bias Evaluation and Assessment Test Suite for Large Language Models
by: Abhishek, Alok, et al.
Published: (2025)
by: Abhishek, Alok, et al.
Published: (2025)
Reasoning Promotes Robustness in Theory of Mind Tasks
by: de Haan, Ian B., et al.
Published: (2026)
by: de Haan, Ian B., et al.
Published: (2026)
Lambek pregroups are Frobenius spiders in preorders
by: Pavlovic, Dusko
Published: (2021)
by: Pavlovic, Dusko
Published: (2021)
Measuring Alignment-Induced Activation Shifts Correctly: A Template-Controlled Difference-in-Differences Protocol
by: Nakamura, Yuki
Published: (2026)
by: Nakamura, Yuki
Published: (2026)
The compact double category $\mathbf{Int}(\mathbf{Poly}_*)$ models control flow and data transformations
by: Kondyrev, Grigory, et al.
Published: (2025)
by: Kondyrev, Grigory, et al.
Published: (2025)
VQToken: Neural Discrete Token Representation Learning for Extreme Token Reduction in Video Large Language Models
by: Zhang, Haichao, et al.
Published: (2025)
by: Zhang, Haichao, et al.
Published: (2025)
The Knesset Corpus: An Annotated Corpus of Hebrew Parliamentary Proceedings
by: Goldin, Gili, et al.
Published: (2024)
by: Goldin, Gili, et al.
Published: (2024)
Math Natural Language Inference: this should be easy!
by: de Paiva, Valeria, et al.
Published: (2025)
by: de Paiva, Valeria, et al.
Published: (2025)
Fast Quiet-STaR: Thinking Without Thought Tokens
by: Huang, Wei, et al.
Published: (2025)
by: Huang, Wei, et al.
Published: (2025)
New Skills or Sharper Primitives? A Probabilistic Perspective on the Emergence of Reasoning in RLVR
by: Wang, Zhilin, et al.
Published: (2026)
by: Wang, Zhilin, et al.
Published: (2026)
d-TreeRPO: Towards More Reliable Policy Optimization for Diffusion Language Models
by: Pan, Leyi, et al.
Published: (2025)
by: Pan, Leyi, et al.
Published: (2025)
Similar Items
-
A Mathematical Theory of Discursive Networks
by: Gutiérrez, Juan B.
Published: (2025) -
On internal categories and crossed objects in the category of monoids
by: Pirashvili, Ilia
Published: (2024) -
Recipient Profiling: Predicting Characteristics from Messages
by: Borquez, Martin, et al.
Published: (2024) -
Data and AI governance: Promoting equity, ethics, and fairness in large language models
by: Abhishek, Alok, et al.
Published: (2025) -
The 99% Success Paradox: When Near-Perfect Retrieval Equals Random Selection
by: Repantis, Vyzantinos, et al.
Published: (2026)