Measurement Risk in Supervised Financial NLP: Rubric and Metric Sensitivity on JF-ICR
Fuente:
arXiv
Saved in:
| Main Authors: | Chang, Sidi, Zhu, Peiying, Chen, Yuxiao, Chai, Rongdong |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Word Importance Explains How Prompts Affect Language Model Outputs
by: Hackmann, Stefan, et al.
Published: (2024)
by: Hackmann, Stefan, et al.
Published: (2024)
Classification of descriptions and summary using multiple passes of statistical and natural language toolkits
by: Banthia, Saumya, et al.
Published: (2020)
by: Banthia, Saumya, et al.
Published: (2020)
Sample-Efficient Language Model for Hinglish Conversational AI
by: Singh, Sakshi, et al.
Published: (2025)
by: Singh, Sakshi, et al.
Published: (2025)
Measuring Robustness of Speech Recognition from MEG Signals Under Distribution Shift
by: Chien, Sheng-You, et al.
Published: (2026)
by: Chien, Sheng-You, et al.
Published: (2026)
RubricRAG: Towards Interpretable and Reliable LLM Evaluation via Domain Knowledge Retrieval for Rubric Generation
by: Dhole, Kaustubh D., et al.
Published: (2026)
by: Dhole, Kaustubh D., et al.
Published: (2026)
Adversarially Probing Cross-Family Sound Symbolism in 27 Languages
by: Sharma, Anika, et al.
Published: (2025)
by: Sharma, Anika, et al.
Published: (2025)
A Multi-Encoder Frozen-Decoder Approach for Fine-Tuning Large Language Models
by: Dhole, Kaustubh D.
Published: (2025)
by: Dhole, Kaustubh D.
Published: (2025)
ValueBlindBench: Agreement-Gated Stress Testing of LLM-Judged Investment Rationales Before Returns Are Observable
by: Chang, Sidi, et al.
Published: (2026)
by: Chang, Sidi, et al.
Published: (2026)
Project Riley: Multimodal Multi-Agent LLM Collaboration with Emotional Reasoning and Voting
by: Ortigoso, Ana Rita, et al.
Published: (2025)
by: Ortigoso, Ana Rita, et al.
Published: (2025)
The Rise of Verbal Tics in Large Language Models: A Systematic Analysis Across Frontier Models
by: Wu, Shuai, et al.
Published: (2026)
by: Wu, Shuai, et al.
Published: (2026)
"The Data Says Otherwise"-Towards Automated Fact-checking and Communication of Data Claims
by: Fu, Yu, et al.
Published: (2024)
by: Fu, Yu, et al.
Published: (2024)
ECLAIR: Enhanced Clarification for Interactive Responses in an Enterprise AI Assistant
by: Murzaku, John, et al.
Published: (2025)
by: Murzaku, John, et al.
Published: (2025)
Quantifying Cognitive Bias Induction in LLM-Generated Content
by: Alessa, Abeer, et al.
Published: (2025)
by: Alessa, Abeer, et al.
Published: (2025)
Human Agency, Causality, and the Human Computer Interface in High-Stakes Artificial Intelligence
by: Hattab, Georges
Published: (2026)
by: Hattab, Georges
Published: (2026)
Distilled HuBERT for Mobile Speech Emotion Recognition: A Cross-Corpus Validation Study
by: Ismail, Saifelden M.
Published: (2025)
by: Ismail, Saifelden M.
Published: (2025)
A Reflective Storytelling Agent for Older Adults: Integrating Argumentation Schemes and Argument Mining in LLM-Based Personalised Narratives
by: Baskar, Jayalakshmi, et al.
Published: (2026)
by: Baskar, Jayalakshmi, et al.
Published: (2026)
The Silicon Mirror: Dynamic Behavioral Gating for Anti-Sycophancy in LLM Agents
by: Shah, Harshee Jignesh
Published: (2026)
by: Shah, Harshee Jignesh
Published: (2026)
Evaluating 5W3H Structured Prompting for Intent Alignment in Human-AI Interaction
by: Gang, Peng
Published: (2026)
by: Gang, Peng
Published: (2026)
Inaccuracy of an E-Dictionary and Its Influence on Chinese Language Users
by: Zhang, Shiyang, et al.
Published: (2025)
by: Zhang, Shiyang, et al.
Published: (2025)
Exploring Mobile Touch Interaction with Large Language Models
by: Zindulka, Tim, et al.
Published: (2025)
by: Zindulka, Tim, et al.
Published: (2025)
Content-Driven Local Response: Supporting Sentence-Level and Message-Level Mobile Email Replies With and Without AI
by: Zindulka, Tim, et al.
Published: (2025)
by: Zindulka, Tim, et al.
Published: (2025)
Writer-Defined AI Personas for On-Demand Feedback Generation
by: Benharrak, Karim, et al.
Published: (2023)
by: Benharrak, Karim, et al.
Published: (2023)
Collage is the New Writing: Exploring the Fragmentation of Text and User Interfaces in AI Tools
by: Buschek, Daniel
Published: (2024)
by: Buschek, Daniel
Published: (2024)
LLMs and people both learn to form conventions -- just not with each other
by: Jones, Cameron R., et al.
Published: (2026)
by: Jones, Cameron R., et al.
Published: (2026)
The AI Memory Gap: Users Misremember What They Created With AI or Without
by: Zindulka, Tim, et al.
Published: (2025)
by: Zindulka, Tim, et al.
Published: (2025)
Real-Time World Crafting: Generating Structured Game Behaviors from Natural Language with Large Language Models
by: Drake, Austin, et al.
Published: (2025)
by: Drake, Austin, et al.
Published: (2025)
Collaborative Document Editing with Multiple Users and AI Agents
by: Lehmann, Florian, et al.
Published: (2025)
by: Lehmann, Florian, et al.
Published: (2025)
The Adaptation Paradox: Agency vs. Mimicry in Companion Chatbots
by: Brandt, T. James, et al.
Published: (2025)
by: Brandt, T. James, et al.
Published: (2025)
CorpusStudio: Surfacing Emergent Patterns in a Corpus of Prior Work while Writing
by: Dang, Hai, et al.
Published: (2025)
by: Dang, Hai, et al.
Published: (2025)
Composable Prompting Workspaces for Creative Writing: Exploration and Iteration Using Dynamic Widgets
by: Amin, Rifat Mehreen, et al.
Published: (2025)
by: Amin, Rifat Mehreen, et al.
Published: (2025)
PromptCanvas: Composable Prompting Workspaces Using Dynamic Widgets for Exploration and Iteration in Creative Writing
by: Amin, Rifat Mehreen, et al.
Published: (2025)
by: Amin, Rifat Mehreen, et al.
Published: (2025)
MM-Conv: A Multimodal Dataset and Benchmark for Context-Aware Grounding in 3D Dialogue
by: Deichler, Anna, et al.
Published: (2026)
by: Deichler, Anna, et al.
Published: (2026)
Multi-Agent Synergy-Driven Iterative Visual Narrative Synthesis
by: Xi, Wang, et al.
Published: (2025)
by: Xi, Wang, et al.
Published: (2025)
Exploring the Effect of Robotic Embodiment and Empathetic Tone of LLMs on Empathy Elicitation
by: Darwesh, Liza, et al.
Published: (2025)
by: Darwesh, Liza, et al.
Published: (2025)
Human-Robot Dialogue Annotation for Multi-Modal Common Ground
by: Bonial, Claire, et al.
Published: (2024)
by: Bonial, Claire, et al.
Published: (2024)
SCOUT: A Situated and Multi-Modal Human-Robot Dialogue Corpus
by: Lukin, Stephanie M., et al.
Published: (2024)
by: Lukin, Stephanie M., et al.
Published: (2024)
Co-Writing with AI, on Human Terms: Aligning Research with User Demands Across the Writing Process
by: Reza, Mohi, et al.
Published: (2025)
by: Reza, Mohi, et al.
Published: (2025)
Agentic AI Translate: An Agentic Translator Prototype for Translation as Communication Design
by: Yamada, Masaru
Published: (2026)
by: Yamada, Masaru
Published: (2026)
Conversation Tree Architecture: A Structured Framework for Context-Aware Multi-Branch LLM Conversations
by: Hemanth, Pranav, et al.
Published: (2026)
by: Hemanth, Pranav, et al.
Published: (2026)
CommentScope: A Comment-Embedded Assisted Reading System for a Long Text
by: Chen, Shuai, et al.
Published: (2025)
by: Chen, Shuai, et al.
Published: (2025)
Similar Items
-
Word Importance Explains How Prompts Affect Language Model Outputs
by: Hackmann, Stefan, et al.
Published: (2024) -
Classification of descriptions and summary using multiple passes of statistical and natural language toolkits
by: Banthia, Saumya, et al.
Published: (2020) -
Sample-Efficient Language Model for Hinglish Conversational AI
by: Singh, Sakshi, et al.
Published: (2025) -
Measuring Robustness of Speech Recognition from MEG Signals Under Distribution Shift
by: Chien, Sheng-You, et al.
Published: (2026) -
RubricRAG: Towards Interpretable and Reliable LLM Evaluation via Domain Knowledge Retrieval for Rubric Generation
by: Dhole, Kaustubh D., et al.
Published: (2026)