Salvato in:
| Autori principali: | Begin, James, Agrawal, Namit, Singh, Eshan, Fu, Yicheng, O'Brien, Sean, Sharma, Vasu, Zhu, Kevin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2502.20405 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts
di: Gupta, Abhay, et al.
Pubblicazione: (2025)
di: Gupta, Abhay, et al.
Pubblicazione: (2025)
Sarc7: Evaluating Sarcasm Detection and Generation with Seven Types and Emotion-Informed Techniques
di: Xiong, Lang, et al.
Pubblicazione: (2025)
di: Xiong, Lang, et al.
Pubblicazione: (2025)
ERGO: Entropy-guided Resetting for Generation Optimization in Multi-turn Language Models
di: Khalid, Haziq Mohammad, et al.
Pubblicazione: (2025)
di: Khalid, Haziq Mohammad, et al.
Pubblicazione: (2025)
WOLF: Werewolf-based Observations for LLM Deception and Falsehoods
di: Agarwal, Mrinal, et al.
Pubblicazione: (2025)
di: Agarwal, Mrinal, et al.
Pubblicazione: (2025)
Reasoning Relay: Evaluating Stability and Interchangeability of Large Language Models in Mathematical Reasoning
di: Lu, Leo, et al.
Pubblicazione: (2025)
di: Lu, Leo, et al.
Pubblicazione: (2025)
SMAGDi: Socratic Multi Agent Interaction Graph Distillation for Efficient High Accuracy Reasoning
di: Aluru, Aayush, et al.
Pubblicazione: (2025)
di: Aluru, Aayush, et al.
Pubblicazione: (2025)
Pruning for Performance: Efficient Idiom and Metaphor Classification in Low-Resource Konkani Using mBERT
di: Do, Timothy, et al.
Pubblicazione: (2025)
di: Do, Timothy, et al.
Pubblicazione: (2025)
Question-Analysis Prompting Improves LLM Performance in Reasoning Tasks
di: Yugeswardeenoo, Dharunish, et al.
Pubblicazione: (2024)
di: Yugeswardeenoo, Dharunish, et al.
Pubblicazione: (2024)
FAIRE: Assessing Racial and Gender Bias in AI-Driven Resume Evaluations
di: Wen, Athena, et al.
Pubblicazione: (2025)
di: Wen, Athena, et al.
Pubblicazione: (2025)
Causal Language Control in Multilingual Transformers via Sparse Feature Steering
di: Chou, Cheng-Ting, et al.
Pubblicazione: (2025)
di: Chou, Cheng-Ting, et al.
Pubblicazione: (2025)
Universal Neurons in GPT-2: Emergence, Persistence, and Functional Impact
di: Nandan, Advey, et al.
Pubblicazione: (2025)
di: Nandan, Advey, et al.
Pubblicazione: (2025)
Adaptive Originality Filtering: Rejection Based Prompting and RiddleScore for Culturally Grounded Multilingual Riddle Generation
di: Le, Duy, et al.
Pubblicazione: (2025)
di: Le, Duy, et al.
Pubblicazione: (2025)
From Directions to Cones: Exploring Multidimensional Representations of Propositional Facts in LLMs
di: Yu, Stanley, et al.
Pubblicazione: (2025)
di: Yu, Stanley, et al.
Pubblicazione: (2025)
Distill CLIP (DCLIP): Enhancing Image-Text Retrieval via Cross-Modal Transformer Distillation
di: Csizmadia, Daniel, et al.
Pubblicazione: (2025)
di: Csizmadia, Daniel, et al.
Pubblicazione: (2025)
TRUTH DECAY: Quantifying Multi-Turn Sycophancy in Language Models
di: Liu, Joshua, et al.
Pubblicazione: (2025)
di: Liu, Joshua, et al.
Pubblicazione: (2025)
From Competition to Coordination: Market Making as a Scalable Framework for Safe and Aligned Multi-Agent LLM Systems
di: Gho, Brendan, et al.
Pubblicazione: (2025)
di: Gho, Brendan, et al.
Pubblicazione: (2025)
The Geometry of Harmfulness in LLMs through Subconcept Probing
di: Shah, McNair, et al.
Pubblicazione: (2025)
di: Shah, McNair, et al.
Pubblicazione: (2025)
Deconstructing Bias: A Multifaceted Framework for Diagnosing Cultural and Compositional Inequities in Text-to-Image Generative Models
di: Said, Muna Numan, et al.
Pubblicazione: (2025)
di: Said, Muna Numan, et al.
Pubblicazione: (2025)
Direct Confidence Alignment: Aligning Verbalized Confidence with Internal Confidence In Large Language Models
di: Zhang, Glenn, et al.
Pubblicazione: (2025)
di: Zhang, Glenn, et al.
Pubblicazione: (2025)
COREVQA: A Crowd Observation and Reasoning Entailment Visual Question Answering Benchmark
di: Chintapatla, Ishant, et al.
Pubblicazione: (2025)
di: Chintapatla, Ishant, et al.
Pubblicazione: (2025)
Rosetta-PL: Propositional Logic as a Benchmark for Large Language Model Reasoning
di: Baek, Shaun, et al.
Pubblicazione: (2025)
di: Baek, Shaun, et al.
Pubblicazione: (2025)
Advancing Uto-Aztecan Language Technologies: A Case Study on the Endangered Comanche Language
di: C, Jesus Alvarez, et al.
Pubblicazione: (2025)
di: C, Jesus Alvarez, et al.
Pubblicazione: (2025)
Interpreting the Latent Structure of Operator Precedence in Language Models
di: Yugeswardeenoo, Dharunish, et al.
Pubblicazione: (2025)
di: Yugeswardeenoo, Dharunish, et al.
Pubblicazione: (2025)
AAVENUE: Detecting LLM Biases on NLU Tasks in AAVE via a Novel Benchmark
di: Gupta, Abhay, et al.
Pubblicazione: (2024)
di: Gupta, Abhay, et al.
Pubblicazione: (2024)
Implementing Open Approaches in the School.
di: O'Brien, E.
Pubblicazione: (1977)
di: O'Brien, E.
Pubblicazione: (1977)
ChunkRAG: Novel LLM-Chunk Filtering Method for RAG Systems
di: Singh, Ishneet Sukhvinder, et al.
Pubblicazione: (2024)
di: Singh, Ishneet Sukhvinder, et al.
Pubblicazione: (2024)
Traversing European Coastlines (TREC) particle count and meterological data from land (2023-2024)
di: O'Brien, James
Pubblicazione: (2026)
di: O'Brien, James
Pubblicazione: (2026)
Error Reflection Prompting: Can Large Language Models Successfully Understand Errors?
di: Li, Jason, et al.
Pubblicazione: (2025)
di: Li, Jason, et al.
Pubblicazione: (2025)
CLEAR: Contrasting Textual Feedback with Experts and Amateurs for Reasoning
di: Rufail, Andrew, et al.
Pubblicazione: (2025)
di: Rufail, Andrew, et al.
Pubblicazione: (2025)
Improving LLM Abilities in Idiomatic Translation
di: Donthi, Sundesh, et al.
Pubblicazione: (2024)
di: Donthi, Sundesh, et al.
Pubblicazione: (2024)
Sistemas de información gerencial / James A. O’Brien, George M. Marakas ; traducción, María Jesús Herrero Díaz, Miguel µngel S nchez Carrión
di: O’Brien, James A
di: O’Brien, James A
Federated Stream-Processing and Latency-Gated Response for Cross-Sector Threat Detection and Collaborative Containment
di: Mohale, Namit
Pubblicazione: (2026)
di: Mohale, Namit
Pubblicazione: (2026)
In-In EFT
di: Mahajan, Namit
Pubblicazione: (2025)
di: Mahajan, Namit
Pubblicazione: (2025)
MALIBU Benchmark: Multi-Agent LLM Implicit Bias Uncovered
di: Mirza, Imran, et al.
Pubblicazione: (2025)
di: Mirza, Imran, et al.
Pubblicazione: (2025)
Ultrafast Superconducting Qubit Readout with the Quarton Coupler
di: Ye, Yufeng, et al.
Pubblicazione: (2024)
di: Ye, Yufeng, et al.
Pubblicazione: (2024)
From Bias to Balance: Detecting Facial Expression Recognition Biases in Large Multimodal Foundation Models
di: Chhua, Kaylee, et al.
Pubblicazione: (2024)
di: Chhua, Kaylee, et al.
Pubblicazione: (2024)
Revolutionizing Early Detection: A Comprehensive Cognitive Screening Tool for ADRD Through Innovative Mobile App Technology
di: Lenora W Smith, et al.
Pubblicazione: (2024)
di: Lenora W Smith, et al.
Pubblicazione: (2024)
The Ethics of AI generated artworks
di: Eshan Divecha
Pubblicazione: (2022)
di: Eshan Divecha
Pubblicazione: (2022)
Using a Capability Approach to Explore How People With Intellectual Disabilities Can Lead Flourishing Lives
di: Sara Ryan, et al.
Pubblicazione: (2024)
di: Sara Ryan, et al.
Pubblicazione: (2024)
A Few Bad Neurons: Isolating and Surgically Correcting Sycophancy
di: O'Brien, Claire, et al.
Pubblicazione: (2026)
di: O'Brien, Claire, et al.
Pubblicazione: (2026)
Documenti analoghi
-
NovelHopQA: Diagnosing Multi-Hop Reasoning Failures in Long Narrative Contexts
di: Gupta, Abhay, et al.
Pubblicazione: (2025) -
Sarc7: Evaluating Sarcasm Detection and Generation with Seven Types and Emotion-Informed Techniques
di: Xiong, Lang, et al.
Pubblicazione: (2025) -
ERGO: Entropy-guided Resetting for Generation Optimization in Multi-turn Language Models
di: Khalid, Haziq Mohammad, et al.
Pubblicazione: (2025) -
WOLF: Werewolf-based Observations for LLM Deception and Falsehoods
di: Agarwal, Mrinal, et al.
Pubblicazione: (2025) -
Reasoning Relay: Evaluating Stability and Interchangeability of Large Language Models in Mathematical Reasoning
di: Lu, Leo, et al.
Pubblicazione: (2025)