TACT: Advancing Complex Aggregative Reasoning with Information Extraction Tools
Fuente:
arXiv
Guardado en:
| Autores principales: | Caciularu, Avi, Jacovi, Alon, Ben-David, Eyal, Goldshtein, Sasha, Schuster, Tal, Herzig, Jonathan, Elidan, Gal, Globerson, Amir |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
DRAGged into Conflicts: Detecting and Addressing Conflicting Sources in Search-Augmented LLMs
por: Cattan, Arie, et al.
Publicado: (2025)
por: Cattan, Arie, et al.
Publicado: (2025)
CoverBench: A Challenging Benchmark for Complex Claim Verification
por: Jacovi, Alon, et al.
Publicado: (2024)
por: Jacovi, Alon, et al.
Publicado: (2024)
Latent Reasoning with Supervised Thinking States
por: Amos, Ido, et al.
Publicado: (2026)
por: Amos, Ido, et al.
Publicado: (2026)
Do LLMs have Consistent Values?
por: Rozen, Naama, et al.
Publicado: (2024)
por: Rozen, Naama, et al.
Publicado: (2024)
Large Language Models for Water Distribution Systems Modeling and Decision-Making
por: Goldshtein, Yinon, et al.
Publicado: (2025)
por: Goldshtein, Yinon, et al.
Publicado: (2025)
DoubleDipper: Improving Long-Context LLMs via Context Recycling
por: Cattan, Arie, et al.
Publicado: (2024)
por: Cattan, Arie, et al.
Publicado: (2024)
Teaching Models to Improve on Tape
por: Bezalel, Liat, et al.
Publicado: (2024)
por: Bezalel, Liat, et al.
Publicado: (2024)
On the Optimization Landscape of Maximum Mean Discrepancy
por: Alon, Itai, et al.
Publicado: (2021)
por: Alon, Itai, et al.
Publicado: (2021)
Fast Inference via Hierarchical Speculative Decoding
por: Mohri, Clara, et al.
Publicado: (2025)
por: Mohri, Clara, et al.
Publicado: (2025)
ConvApparel: A Benchmark Dataset and Validation Framework for User Simulators in Conversational Recommenders
por: Meshi, Ofer, et al.
Publicado: (2026)
por: Meshi, Ofer, et al.
Publicado: (2026)
A Chain-of-Thought Is as Strong as Its Weakest Link: A Benchmark for Verifiers of Reasoning Chains
por: Jacovi, Alon, et al.
Publicado: (2024)
por: Jacovi, Alon, et al.
Publicado: (2024)
Unpacking Tokenization: Evaluating Text Compression and its Correlation with Model Performance
por: Goldman, Omer, et al.
Publicado: (2024)
por: Goldman, Omer, et al.
Publicado: (2024)
Does Fine-Tuning LLMs on New Knowledge Encourage Hallucinations?
por: Gekhman, Zorik, et al.
Publicado: (2024)
por: Gekhman, Zorik, et al.
Publicado: (2024)
SimpleQA Verified: A Reliable Factuality Benchmark to Measure Parametric Knowledge
por: Haas, Lukas, et al.
Publicado: (2025)
por: Haas, Lukas, et al.
Publicado: (2025)
Span-Aggregatable, Contextualized Word Embeddings for Effective Phrase Mining
por: Orbach, Eyal, et al.
Publicado: (2024)
por: Orbach, Eyal, et al.
Publicado: (2024)
Dont Add, dont Miss: Effective Content Preserving Generation from Pre-Selected Text Spans
por: Slobodkin, Aviv, et al.
Publicado: (2023)
por: Slobodkin, Aviv, et al.
Publicado: (2023)
MetaFaith: Faithful Natural Language Uncertainty Expression in LLMs
por: Liu, Gabrielle Kaili-May, et al.
Publicado: (2025)
por: Liu, Gabrielle Kaili-May, et al.
Publicado: (2025)
Unpacking Human-AI Interaction in Safety-Critical Industries: A Systematic Literature Review
por: Bach, Tita A., et al.
Publicado: (2023)
por: Bach, Tita A., et al.
Publicado: (2023)
ReliableEval: A Recipe for Stochastic LLM Evaluation via Method of Moments
por: Lior, Gili, et al.
Publicado: (2025)
por: Lior, Gili, et al.
Publicado: (2025)
Patchscopes: A Unifying Framework for Inspecting Hidden Representations of Language Models
por: Ghandeharioun, Asma, et al.
Publicado: (2024)
por: Ghandeharioun, Asma, et al.
Publicado: (2024)
Agency independence and credibility in primary bond markets
por: Tal Sadeh, et al.
Publicado: (2024)
por: Tal Sadeh, et al.
Publicado: (2024)
Temporal Aggregation for the Synthetic Control Method
por: Sun, Liyang, et al.
Publicado: (2024)
por: Sun, Liyang, et al.
Publicado: (2024)
ConSim: Measuring Concept-Based Explanations' Effectiveness with Automated Simulatability
por: Poché, Antonin, et al.
Publicado: (2025)
por: Poché, Antonin, et al.
Publicado: (2025)
Fooling Algorithms in Non-Stationary Bandits using Belief Inertia
por: Mendelson, Gal, et al.
Publicado: (2025)
por: Mendelson, Gal, et al.
Publicado: (2025)
Diverse Inference and Verification for Advanced Reasoning
por: Drori, Iddo, et al.
Publicado: (2025)
por: Drori, Iddo, et al.
Publicado: (2025)
How Does Overparameterization Affect Machine Unlearning of Deep Neural Networks?
por: Alon, Gal, et al.
Publicado: (2025)
por: Alon, Gal, et al.
Publicado: (2025)
Is It Really Long Context if All You Need Is Retrieval? Towards Genuinely Difficult Long Context NLP
por: Goldman, Omer, et al.
Publicado: (2024)
por: Goldman, Omer, et al.
Publicado: (2024)
Efficient Time Series Forecasting via Hyper-Complex Models and Frequency Aggregation
por: Yakir, Eyal, et al.
Publicado: (2025)
por: Yakir, Eyal, et al.
Publicado: (2025)
TACT: Mitigating Overthinking and Overacting in Coding Agents via Activation Steering
por: Sui, Yuan, et al.
Publicado: (2026)
por: Sui, Yuan, et al.
Publicado: (2026)
SEAM: A Stochastic Benchmark for Multi-Document Tasks
por: Lior, Gili, et al.
Publicado: (2024)
por: Lior, Gili, et al.
Publicado: (2024)
MDCure: A Scalable Pipeline for Multi-Document Instruction-Following
por: Liu, Gabrielle Kaili-May, et al.
Publicado: (2024)
por: Liu, Gabrielle Kaili-May, et al.
Publicado: (2024)
Identifying User Goals from UI Trajectories
por: Berkovitch, Omri, et al.
Publicado: (2024)
por: Berkovitch, Omri, et al.
Publicado: (2024)
Random Walks in Self-supervised Learning for Triangular Meshes
por: Yefet, Gal, et al.
Publicado: (2025)
por: Yefet, Gal, et al.
Publicado: (2025)
Empty Shelves or Lost Keys? Recall Is the Bottleneck for Parametric Factuality
por: Calderon, Nitay, et al.
Publicado: (2026)
por: Calderon, Nitay, et al.
Publicado: (2026)
Linearized analysis of dissipative Two Axis Counter Twisting (TACT) squeezing for Metrology
por: Goldstein, Garry
Publicado: (2024)
por: Goldstein, Garry
Publicado: (2024)
Inside-Out: Hidden Factual Knowledge in LLMs
por: Gekhman, Zorik, et al.
Publicado: (2025)
por: Gekhman, Zorik, et al.
Publicado: (2025)
The Complexity of Aggregates over Extractions by Regular Expressions
por: Doleschal, Johannes, et al.
Publicado: (2020)
por: Doleschal, Johannes, et al.
Publicado: (2020)
Finding Visual Task Vectors
por: Hojel, Alberto, et al.
Publicado: (2024)
por: Hojel, Alberto, et al.
Publicado: (2024)
Cost-Sensitive Neighborhood Aggregation for Heterophilous Graphs: When Does Per-Edge Routing Help?
por: Weiss, Eyal
Publicado: (2026)
por: Weiss, Eyal
Publicado: (2026)
TabAgent: A Framework for Replacing Agentic Generative Components with Tabular-Textual Classifiers
por: Levy, Ido, et al.
Publicado: (2026)
por: Levy, Ido, et al.
Publicado: (2026)
Ejemplares similares
-
DRAGged into Conflicts: Detecting and Addressing Conflicting Sources in Search-Augmented LLMs
por: Cattan, Arie, et al.
Publicado: (2025) -
CoverBench: A Challenging Benchmark for Complex Claim Verification
por: Jacovi, Alon, et al.
Publicado: (2024) -
Latent Reasoning with Supervised Thinking States
por: Amos, Ido, et al.
Publicado: (2026) -
Do LLMs have Consistent Values?
por: Rozen, Naama, et al.
Publicado: (2024) -
Large Language Models for Water Distribution Systems Modeling and Decision-Making
por: Goldshtein, Yinon, et al.
Publicado: (2025)