Salvato in:
| Autori principali: | Said, Muna Numan, Zaidi, Aarib, Usman, Rabia, Okon, Sonia, Medepalli, Praneeth, Zhu, Kevin, Sharma, Vasu, O'Brien, Sean |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2505.01430 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts
di: Gupta, Abhay, et al.
Pubblicazione: (2025)
di: Gupta, Abhay, et al.
Pubblicazione: (2025)
Adaptive Originality Filtering: Rejection Based Prompting and RiddleScore for Culturally Grounded Multilingual Riddle Generation
di: Le, Duy, et al.
Pubblicazione: (2025)
di: Le, Duy, et al.
Pubblicazione: (2025)
Distill CLIP (DCLIP): Enhancing Image-Text Retrieval via Cross-Modal Transformer Distillation
di: Csizmadia, Daniel, et al.
Pubblicazione: (2025)
di: Csizmadia, Daniel, et al.
Pubblicazione: (2025)
Reasoning Relay: Evaluating Stability and Interchangeability of Large Language Models in Mathematical Reasoning
di: Lu, Leo, et al.
Pubblicazione: (2025)
di: Lu, Leo, et al.
Pubblicazione: (2025)
SMAGDi: Socratic Multi Agent Interaction Graph Distillation for Efficient High Accuracy Reasoning
di: Aluru, Aayush, et al.
Pubblicazione: (2025)
di: Aluru, Aayush, et al.
Pubblicazione: (2025)
Sarc7: Evaluating Sarcasm Detection and Generation with Seven Types and Emotion-Informed Techniques
di: Xiong, Lang, et al.
Pubblicazione: (2025)
di: Xiong, Lang, et al.
Pubblicazione: (2025)
Pause-Tuning for Long-Context Comprehension: A Lightweight Approach to LLM Attention Recalibration
di: Begin, James, et al.
Pubblicazione: (2025)
di: Begin, James, et al.
Pubblicazione: (2025)
Pruning for Performance: Efficient Idiom and Metaphor Classification in Low-Resource Konkani Using mBERT
di: Do, Timothy, et al.
Pubblicazione: (2025)
di: Do, Timothy, et al.
Pubblicazione: (2025)
Causal Language Control in Multilingual Transformers via Sparse Feature Steering
di: Chou, Cheng-Ting, et al.
Pubblicazione: (2025)
di: Chou, Cheng-Ting, et al.
Pubblicazione: (2025)
ERGO: Entropy-guided Resetting for Generation Optimization in Multi-turn Language Models
di: Khalid, Haziq Mohammad, et al.
Pubblicazione: (2025)
di: Khalid, Haziq Mohammad, et al.
Pubblicazione: (2025)
Universal Neurons in GPT-2: Emergence, Persistence, and Functional Impact
di: Nandan, Advey, et al.
Pubblicazione: (2025)
di: Nandan, Advey, et al.
Pubblicazione: (2025)
From Directions to Cones: Exploring Multidimensional Representations of Propositional Facts in LLMs
di: Yu, Stanley, et al.
Pubblicazione: (2025)
di: Yu, Stanley, et al.
Pubblicazione: (2025)
WOLF: Werewolf-based Observations for LLM Deception and Falsehoods
di: Agarwal, Mrinal, et al.
Pubblicazione: (2025)
di: Agarwal, Mrinal, et al.
Pubblicazione: (2025)
TRUTH DECAY: Quantifying Multi-Turn Sycophancy in Language Models
di: Liu, Joshua, et al.
Pubblicazione: (2025)
di: Liu, Joshua, et al.
Pubblicazione: (2025)
The illusion of academic freedom and the promise of the undercommons
di: Zareen Zaidi, et al.
Pubblicazione: (2026)
di: Zareen Zaidi, et al.
Pubblicazione: (2026)
The Geometry of Harmfulness in LLMs through Subconcept Probing
di: Shah, McNair, et al.
Pubblicazione: (2025)
di: Shah, McNair, et al.
Pubblicazione: (2025)
From Bias to Balance: Detecting Facial Expression Recognition Biases in Large Multimodal Foundation Models
di: Chhua, Kaylee, et al.
Pubblicazione: (2024)
di: Chhua, Kaylee, et al.
Pubblicazione: (2024)
COREVQA: A Crowd Observation and Reasoning Entailment Visual Question Answering Benchmark
di: Chintapatla, Ishant, et al.
Pubblicazione: (2025)
di: Chintapatla, Ishant, et al.
Pubblicazione: (2025)
Rosetta-PL: Propositional Logic as a Benchmark for Large Language Model Reasoning
di: Baek, Shaun, et al.
Pubblicazione: (2025)
di: Baek, Shaun, et al.
Pubblicazione: (2025)
Advancing Uto-Aztecan Language Technologies: A Case Study on the Endangered Comanche Language
di: C, Jesus Alvarez, et al.
Pubblicazione: (2025)
di: C, Jesus Alvarez, et al.
Pubblicazione: (2025)
FAIRE: Assessing Racial and Gender Bias in AI-Driven Resume Evaluations
di: Wen, Athena, et al.
Pubblicazione: (2025)
di: Wen, Athena, et al.
Pubblicazione: (2025)
Interpreting the Latent Structure of Operator Precedence in Language Models
di: Yugeswardeenoo, Dharunish, et al.
Pubblicazione: (2025)
di: Yugeswardeenoo, Dharunish, et al.
Pubblicazione: (2025)
Question-Analysis Prompting Improves LLM Performance in Reasoning Tasks
di: Yugeswardeenoo, Dharunish, et al.
Pubblicazione: (2024)
di: Yugeswardeenoo, Dharunish, et al.
Pubblicazione: (2024)
MALIBU Benchmark: Multi-Agent LLM Implicit Bias Uncovered
di: Mirza, Imran, et al.
Pubblicazione: (2025)
di: Mirza, Imran, et al.
Pubblicazione: (2025)
Influence of Rhizophagus irregularis Inoculation on Salt Tolerance in Cucurbita maxima Duch.
di: Okon, Okon Godwin, et al.
Pubblicazione: (2018)
di: Okon, Okon Godwin, et al.
Pubblicazione: (2018)
Probing Audio-Generation Capabilities of Text-Based Language Models
di: Anbazhagan, Arjun Prasaath, et al.
Pubblicazione: (2025)
di: Anbazhagan, Arjun Prasaath, et al.
Pubblicazione: (2025)
Deconstructing FastText
di: Majumdar, Partha
Pubblicazione: (2026)
di: Majumdar, Partha
Pubblicazione: (2026)
Impact of Green Knowledge Sharing on the Organizational Performance of SMEs : The Mediating Role of Green Organizational Culture and Technological Innovation
di: Fernando Almeida, et al.
Pubblicazione: (2026)
di: Fernando Almeida, et al.
Pubblicazione: (2026)
Error Reflection Prompting: Can Large Language Models Successfully Understand Errors?
di: Li, Jason, et al.
Pubblicazione: (2025)
di: Li, Jason, et al.
Pubblicazione: (2025)
CLEAR: Contrasting Textual Feedback with Experts and Amateurs for Reasoning
di: Rufail, Andrew, et al.
Pubblicazione: (2025)
di: Rufail, Andrew, et al.
Pubblicazione: (2025)
Comparison of Radiofrequency Microneedling and Ultrasound Delivery of Plant‐Based Derived Secretory Factor (CFa1) Hair Serum for the Cosmetic Improvement of Androgenetic Alopecia
di: Lauren S. Mohan, et al.
Pubblicazione: (2026)
di: Lauren S. Mohan, et al.
Pubblicazione: (2026)
Rewrite-to-Rank: Optimizing Ad Visibility via Retrieval-Aware Text Rewriting
di: Ho, Chloe, et al.
Pubblicazione: (2025)
di: Ho, Chloe, et al.
Pubblicazione: (2025)
Flavour Deconstructing the Composite Higgs
di: Covone, Sebastiano, et al.
Pubblicazione: (2024)
di: Covone, Sebastiano, et al.
Pubblicazione: (2024)
Vopěnka's Principle, Maximum Deconstructibility, and singly-generated torsion classes
di: Cox, Sean
Pubblicazione: (2024)
di: Cox, Sean
Pubblicazione: (2024)
AAVENUE: Detecting LLM Biases on NLU Tasks in AAVE via a Novel Benchmark
di: Gupta, Abhay, et al.
Pubblicazione: (2024)
di: Gupta, Abhay, et al.
Pubblicazione: (2024)
Ultrafast Superconducting Qubit Readout with the Quarton Coupler
di: Ye, Yufeng, et al.
Pubblicazione: (2024)
di: Ye, Yufeng, et al.
Pubblicazione: (2024)
El problema de la medición en mecánica cuántica
di: E. Okon
Pubblicazione: (2014)
di: E. Okon
Pubblicazione: (2014)
Reassessing the strength of a class of Wigner's friend no-go theorems
di: Okon, E.
Pubblicazione: (2022)
di: Okon, E.
Pubblicazione: (2022)
Deconstructive Composite Dark Matter Detection
di: Boukhtouchen, Yilda, et al.
Pubblicazione: (2025)
di: Boukhtouchen, Yilda, et al.
Pubblicazione: (2025)
Encoding Inequity: Examining Demographic Bias in LLM-Driven Robot Caregiving
di: Korpan, Raj
Pubblicazione: (2025)
di: Korpan, Raj
Pubblicazione: (2025)
Documenti analoghi
-
NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts
di: Gupta, Abhay, et al.
Pubblicazione: (2025) -
Adaptive Originality Filtering: Rejection Based Prompting and RiddleScore for Culturally Grounded Multilingual Riddle Generation
di: Le, Duy, et al.
Pubblicazione: (2025) -
Distill CLIP (DCLIP): Enhancing Image-Text Retrieval via Cross-Modal Transformer Distillation
di: Csizmadia, Daniel, et al.
Pubblicazione: (2025) -
Reasoning Relay: Evaluating Stability and Interchangeability of Large Language Models in Mathematical Reasoning
di: Lu, Leo, et al.
Pubblicazione: (2025) -
SMAGDi: Socratic Multi Agent Interaction Graph Distillation for Efficient High Accuracy Reasoning
di: Aluru, Aayush, et al.
Pubblicazione: (2025)