A dataset of questions on decision-theoretic reasoning in Newcomb-like problems
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Oesterheld, Caspar, Cooper, Emery, Kodama, Miles, Nguyen, Linh Chi, Perez, Ethan |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
GanitBench: A bi-lingual benchmark for evaluating mathematical reasoning in Vision Language Models
par: Bandooni, Ashutosh, et autres
Publié: (2025)
par: Bandooni, Ashutosh, et autres
Publié: (2025)
RomanLens: The Role Of Latent Romanization In Multilinguality In LLMs
par: Saji, Alan, et autres
Publié: (2025)
par: Saji, Alan, et autres
Publié: (2025)
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
par: Peters, Sydney, et autres
Publié: (2025)
par: Peters, Sydney, et autres
Publié: (2025)
LASTIST: LArge-Scale Target-Independent STance dataset
par: Kim, DongJae, et autres
Publié: (2025)
par: Kim, DongJae, et autres
Publié: (2025)
Transforming Dutch: Debiasing Dutch Coreference Resolution Systems for Non-binary Pronouns
par: van Boven, Goya, et autres
Publié: (2024)
par: van Boven, Goya, et autres
Publié: (2024)
Low-Resource Court Judgment Summarization for Common Law Systems
par: Liu, Shuaiqi, et autres
Publié: (2024)
par: Liu, Shuaiqi, et autres
Publié: (2024)
Evaluating the Efficacy of Hybrid Deep Learning Models in Distinguishing AI-Generated Text
par: Oketunji, Abiodun Finbarrs
Publié: (2023)
par: Oketunji, Abiodun Finbarrs
Publié: (2023)
ChatGPT4PCG Competition: Character-like Level Generation for Science Birds
par: Taveekitworachai, Pittawat, et autres
Publié: (2023)
par: Taveekitworachai, Pittawat, et autres
Publié: (2023)
Verbosity Tradeoffs and the Impact of Scale on the Faithfulness of LLM Self-Explanations
par: Siegel, Noah Y., et autres
Publié: (2025)
par: Siegel, Noah Y., et autres
Publié: (2025)
Bielik-Minitron-7B: Compressing Large Language Models via Structured Pruning and Knowledge Distillation for the Polish Language
par: Kinas, Remigiusz, et autres
Publié: (2026)
par: Kinas, Remigiusz, et autres
Publié: (2026)
Neural Machine Translation for Malayalam Paraphrase Generation
par: Varghese, Christeena, et autres
Publié: (2024)
par: Varghese, Christeena, et autres
Publié: (2024)
Is Our Chatbot Telling Lies? Assessing Correctness of an LLM-based Dutch Support Chatbot
par: Lassche, Herman, et autres
Publié: (2024)
par: Lassche, Herman, et autres
Publié: (2024)
Predictive Simultaneous Interpretation: Harnessing Large Language Models for Democratizing Real-Time Multilingual Communication
par: Iida, Kurando, et autres
Publié: (2024)
par: Iida, Kurando, et autres
Publié: (2024)
Exploring RWKV for Sentence Embeddings: Layer-wise Analysis and Baseline Comparison for Semantic Similarity
par: Pan, Xinghan
Publié: (2025)
par: Pan, Xinghan
Publié: (2025)
UA-Legal-Bench: A Benchmark for Evaluating Large Language Models on Ukrainian Legal Reasoning
par: Ovcharov, Volodymyr
Publié: (2026)
par: Ovcharov, Volodymyr
Publié: (2026)
KV-Fold: One-Step KV-Cache Recurrence for Long-Context Inference
par: Nadali, Alireza, et autres
Publié: (2026)
par: Nadali, Alireza, et autres
Publié: (2026)
Latent Personality Alignment: Improving Harmlessness Without Mentioning Harms
par: Le, Linh, et autres
Publié: (2026)
par: Le, Linh, et autres
Publié: (2026)
The Unified Cognitive Consciousness Theory for Language Models: Anchoring Semantics, Thresholds of Activation, and Emergent Reasoning
par: Chang, Edward Y., et autres
Publié: (2025)
par: Chang, Edward Y., et autres
Publié: (2025)
Automated Web Application Testing: End-to-End Test Case Generation with Large Language Models and Screen Transition Graphs
par: Le, Nguyen-Khang, et autres
Publié: (2025)
par: Le, Nguyen-Khang, et autres
Publié: (2025)
Robustness of Large Language Models to Perturbations in Text
par: Singh, Ayush, et autres
Publié: (2024)
par: Singh, Ayush, et autres
Publié: (2024)
Large Language Model (LLM) Bias Index -- LLMBI
par: Oketunji, Abiodun Finbarrs, et autres
Publié: (2023)
par: Oketunji, Abiodun Finbarrs, et autres
Publié: (2023)
Evaluating an evidence-guided reinforcement learning framework in aligning light-parameter large language models with decision-making cognition in psychiatric clinical reasoning
par: Lin, Xinxin, et autres
Publié: (2026)
par: Lin, Xinxin, et autres
Publié: (2026)
Beyond Prefixes: Graph-as-Memory Cross-Attention for Knowledge Graph Completion with Large Language Models
par: Liu, Ruitong, et autres
Publié: (2025)
par: Liu, Ruitong, et autres
Publié: (2025)
LLMs as Signal Detectors: Sensitivity, Bias, and the Temperature-Criterion Analogy
par: Cacioli, Jon-Paul
Publié: (2026)
par: Cacioli, Jon-Paul
Publié: (2026)
The Dual-Route Model of Induction
par: Feucht, Sheridan, et autres
Publié: (2025)
par: Feucht, Sheridan, et autres
Publié: (2025)
chDzDT: Word-level morphology-aware language model for Algerian social media text
par: Aries, Abdelkrime
Publié: (2025)
par: Aries, Abdelkrime
Publié: (2025)
Sticker-TTS: Learn to Utilize Historical Experience with a Sticker-driven Test-Time Scaling Framework
par: Chen, Jie, et autres
Publié: (2025)
par: Chen, Jie, et autres
Publié: (2025)
UNO-Bench: A Unified Benchmark for Exploring the Compositional Law Between Uni-modal and Omni-modal in Omni Models
par: Chen, Chen, et autres
Publié: (2025)
par: Chen, Chen, et autres
Publié: (2025)
UniHetero: Could Generation Enhance Understanding for Vision-Language-Model at Large Data Scale?
par: Chen, Fengjiao, et autres
Publié: (2025)
par: Chen, Fengjiao, et autres
Publié: (2025)
Context Is What You Need: The Maximum Effective Context Window for Real World Limits of LLMs
par: Paulsen, Norman
Publié: (2025)
par: Paulsen, Norman
Publié: (2025)
Prompt-Time Symbolic Knowledge Capture with Large Language Models
par: Çöplü, Tolga, et autres
Publié: (2024)
par: Çöplü, Tolga, et autres
Publié: (2024)
DQA: Diagnostic Question Answering for IT Support
par: Kapoor, Vishaal, et autres
Publié: (2026)
par: Kapoor, Vishaal, et autres
Publié: (2026)
Advancing Polish Language Modeling through Tokenizer Optimization in the Bielik v3 7B and 11B Series
par: Ociepa, Krzysztof, et autres
Publié: (2026)
par: Ociepa, Krzysztof, et autres
Publié: (2026)
Artificial Phantasia: Emergent Mental Imagery in Large Language Models
par: McCarty, Morgan, et autres
Publié: (2025)
par: McCarty, Morgan, et autres
Publié: (2025)
Identifying Intensity of the Structure and Content in Tweets and the Discriminative Power of Attributes in Context with Referential Translation Machines
par: Biçici, Ergun
Publié: (2024)
par: Biçici, Ergun
Publié: (2024)
TALE: A Tool-Augmented Framework for Reference-Free Evaluation of Large Language Models
par: Badshah, Sher, et autres
Publié: (2025)
par: Badshah, Sher, et autres
Publié: (2025)
Separating Constraint Compliance from Semantic Accuracy: A Novel Benchmark for Evaluating Instruction-Following Under Compression
par: Baxi, Rahul
Publié: (2025)
par: Baxi, Rahul
Publié: (2025)
Search-R3: Unifying Reasoning and Embedding in Large Language Models
par: Gui, Yuntao, et autres
Publié: (2025)
par: Gui, Yuntao, et autres
Publié: (2025)
Latent Planning Emerges with Scale
par: Hanna, Michael, et autres
Publié: (2026)
par: Hanna, Michael, et autres
Publié: (2026)
AEyeDE: An Attention-Based Attribution Framework for AI-Generated Text Detection
par: Nourbakhsh, Aria, et autres
Publié: (2026)
par: Nourbakhsh, Aria, et autres
Publié: (2026)
Documents similaires
-
GanitBench: A bi-lingual benchmark for evaluating mathematical reasoning in Vision Language Models
par: Bandooni, Ashutosh, et autres
Publié: (2025) -
RomanLens: The Role Of Latent Romanization In Multilinguality In LLMs
par: Saji, Alan, et autres
Publié: (2025) -
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
par: Peters, Sydney, et autres
Publié: (2025) -
LASTIST: LArge-Scale Target-Independent STance dataset
par: Kim, DongJae, et autres
Publié: (2025) -
Transforming Dutch: Debiasing Dutch Coreference Resolution Systems for Non-binary Pronouns
par: van Boven, Goya, et autres
Publié: (2024)