Salvato in:
| Autori principali: | Shlomi, Eliezer, Levy, Ido, Shapira, Eilam, Katz, Michael, Uziel, Guy, Shlomov, Segev, Mashkif, Nir, Reichart, Roi, Keren, Sarah |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2602.04557 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
TabAgent: A Framework for Replacing Agentic Generative Components with Tabular-Textual Classifiers
di: Levy, Ido, et al.
Pubblicazione: (2026)
di: Levy, Ido, et al.
Pubblicazione: (2026)
TabSTAR: A Tabular Foundation Model for Tabular Data with Text Fields
di: Arazi, Alan, et al.
Pubblicazione: (2025)
di: Arazi, Alan, et al.
Pubblicazione: (2025)
The Poisoned Apple Effect: Strategic Manipulation of Mediated Markets via Technology Expansion of AI Agents
di: Shapira, Eilam, et al.
Pubblicazione: (2026)
di: Shapira, Eilam, et al.
Pubblicazione: (2026)
Predicting Decisions of AI Agents from Limited Interaction through Text-Tabular Modeling
di: Shapira, Eilam, et al.
Pubblicazione: (2026)
di: Shapira, Eilam, et al.
Pubblicazione: (2026)
Can LLMs Replace Economic Choice Prediction Labs? The Case of Language-based Persuasion Games
di: Shapira, Eilam, et al.
Pubblicazione: (2024)
di: Shapira, Eilam, et al.
Pubblicazione: (2024)
Donors and Recipients: On Asymmetric Transfer Across Tasks and Languages with Parameter-Efficient Fine-Tuning
di: Dymkiewicz, Kajetan, et al.
Pubblicazione: (2025)
di: Dymkiewicz, Kajetan, et al.
Pubblicazione: (2025)
From Benchmarks to Business Impact: Deploying IBM Generalist Agent in Enterprise Production
di: Shlomov, Segev, et al.
Pubblicazione: (2025)
di: Shlomov, Segev, et al.
Pubblicazione: (2025)
Towards Enterprise-Ready Computer Using Generalist Agent
di: Marreed, Sami, et al.
Pubblicazione: (2025)
di: Marreed, Sami, et al.
Pubblicazione: (2025)
GLEE: A Unified Framework and Benchmark for Language-based Economic Environments
di: Shapira, Eilam, et al.
Pubblicazione: (2024)
di: Shapira, Eilam, et al.
Pubblicazione: (2024)
SNAP: Semantic Stories for Next Activity Prediction
di: Oved, Alon, et al.
Pubblicazione: (2024)
di: Oved, Alon, et al.
Pubblicazione: (2024)
Human Choice Prediction in Language-based Persuasion Games: Simulation-based Off-Policy Evaluation
di: Shapira, Eilam, et al.
Pubblicazione: (2023)
di: Shapira, Eilam, et al.
Pubblicazione: (2023)
What's the Plan? Evaluating and Developing Planning-Aware Techniques for Language Models
di: Hirsch, Eran, et al.
Pubblicazione: (2024)
di: Hirsch, Eran, et al.
Pubblicazione: (2024)
Governance by Construction for Generalist Agents
di: Shlomov, Segev, et al.
Pubblicazione: (2026)
di: Shlomov, Segev, et al.
Pubblicazione: (2026)
Multi-Review Fusion-in-Context
di: Slobodkin, Aviv, et al.
Pubblicazione: (2024)
di: Slobodkin, Aviv, et al.
Pubblicazione: (2024)
MulTaBench: Benchmarking Multimodal Tabular Learning with Text and Image
di: Arazi, Alan, et al.
Pubblicazione: (2026)
di: Arazi, Alan, et al.
Pubblicazione: (2026)
From Grounding to Planning: Benchmarking Bottlenecks in Web Agents
di: Shlomov, Segev, et al.
Pubblicazione: (2024)
di: Shlomov, Segev, et al.
Pubblicazione: (2024)
Text2Model: Text-based Model Induction for Zero-shot Image Classification
di: Amosy, Ohad, et al.
Pubblicazione: (2022)
di: Amosy, Ohad, et al.
Pubblicazione: (2022)
On Behalf of the Stakeholders: Trends in NLP Model Interpretability in the Era of LLMs
di: Calderon, Nitay, et al.
Pubblicazione: (2024)
di: Calderon, Nitay, et al.
Pubblicazione: (2024)
A Systematic Review of NLP for Dementia -- Tasks, Datasets and Opportunities
di: Peled-Cohen, Lotem, et al.
Pubblicazione: (2024)
di: Peled-Cohen, Lotem, et al.
Pubblicazione: (2024)
An Information-Theoretic Approach to Identifying Formulaic Clusters in Textual Data
di: Yoffe, Gideon, et al.
Pubblicazione: (2025)
di: Yoffe, Gideon, et al.
Pubblicazione: (2025)
The Alternative Annotator Test for LLM-as-a-Judge: How to Statistically Justify Replacing Human Annotators with LLMs
di: Calderon, Nitay, et al.
Pubblicazione: (2025)
di: Calderon, Nitay, et al.
Pubblicazione: (2025)
Multi-Domain Explainability of Preferences
di: Calderon, Nitay, et al.
Pubblicazione: (2025)
di: Calderon, Nitay, et al.
Pubblicazione: (2025)
CRISP: Complex Reasoning with Interpretable Step-based Plans
di: Vetzler, Matan, et al.
Pubblicazione: (2025)
di: Vetzler, Matan, et al.
Pubblicazione: (2025)
A Unifying Scheme for Extractive Content Selection Tasks
di: Amar, Shmuel, et al.
Pubblicazione: (2025)
di: Amar, Shmuel, et al.
Pubblicazione: (2025)
Motivation in Large Language Models
di: Nahum, Omer, et al.
Pubblicazione: (2026)
di: Nahum, Omer, et al.
Pubblicazione: (2026)
The Power of Summary-Source Alignments
di: Ernst, Ori, et al.
Pubblicazione: (2024)
di: Ernst, Ori, et al.
Pubblicazione: (2024)
DeLeaker: Dynamic Inference-Time Reweighting For Semantic Leakage Mitigation in Text-to-Image Models
di: Ventura, Mor, et al.
Pubblicazione: (2025)
di: Ventura, Mor, et al.
Pubblicazione: (2025)
ST-WebAgentBench: A Benchmark for Evaluating Safety and Trustworthiness in Web Agents
di: Levy, Ido, et al.
Pubblicazione: (2024)
di: Levy, Ido, et al.
Pubblicazione: (2024)
AgentFixer: From Failure Detection to Fix Recommendations in LLM Agentic Systems
di: Mulian, Hadar, et al.
Pubblicazione: (2026)
di: Mulian, Hadar, et al.
Pubblicazione: (2026)
LIBERTy: A Causal Framework for Benchmarking Concept-Based Explanations of LLMs with Structural Counterfactuals
di: Toker, Gilat, et al.
Pubblicazione: (2026)
di: Toker, Gilat, et al.
Pubblicazione: (2026)
Systematic Biases in LLM Simulations of Debates
di: Taubenfeld, Amir, et al.
Pubblicazione: (2024)
di: Taubenfeld, Amir, et al.
Pubblicazione: (2024)
Consensus or Conflict? Fine-Grained Evaluation of Conflicting Answers in Question-Answering
di: Nachshoni, Eviatar, et al.
Pubblicazione: (2025)
di: Nachshoni, Eviatar, et al.
Pubblicazione: (2025)
Multilinguality at the Edge: Developing Language Models for the Global South
di: Miranda, Lester James V., et al.
Pubblicazione: (2026)
di: Miranda, Lester James V., et al.
Pubblicazione: (2026)
Are LLMs Better than Reported? Detecting Label Errors and Mitigating Their Effect on Model Performance
di: Nahum, Omer, et al.
Pubblicazione: (2024)
di: Nahum, Omer, et al.
Pubblicazione: (2024)
EncodeRec: An Embedding Backbone for Recommendation Systems
di: Hadad, Guy, et al.
Pubblicazione: (2026)
di: Hadad, Guy, et al.
Pubblicazione: (2026)
Can LLMs Learn Macroeconomic Narratives from Social Media?
di: Gueta, Almog, et al.
Pubblicazione: (2024)
di: Gueta, Almog, et al.
Pubblicazione: (2024)
AdaptiVocab: Enhancing LLM Efficiency in Focused Domains through Lightweight Vocabulary Adaptation
di: Nakash, Itay, et al.
Pubblicazione: (2025)
di: Nakash, Itay, et al.
Pubblicazione: (2025)
Fine-Grained Detection of Context-Grounded Hallucinations Using LLMs
di: Peisakhovsky, Yehonatan, et al.
Pubblicazione: (2025)
di: Peisakhovsky, Yehonatan, et al.
Pubblicazione: (2025)
Navigating Cultural Chasms: Exploring and Unlocking the Cultural POV of Text-To-Image Models
di: Ventura, Mor, et al.
Pubblicazione: (2023)
di: Ventura, Mor, et al.
Pubblicazione: (2023)
NL-Eye: Abductive NLI for Images
di: Ventura, Mor, et al.
Pubblicazione: (2024)
di: Ventura, Mor, et al.
Pubblicazione: (2024)
Documenti analoghi
-
TabAgent: A Framework for Replacing Agentic Generative Components with Tabular-Textual Classifiers
di: Levy, Ido, et al.
Pubblicazione: (2026) -
TabSTAR: A Tabular Foundation Model for Tabular Data with Text Fields
di: Arazi, Alan, et al.
Pubblicazione: (2025) -
The Poisoned Apple Effect: Strategic Manipulation of Mediated Markets via Technology Expansion of AI Agents
di: Shapira, Eilam, et al.
Pubblicazione: (2026) -
Predicting Decisions of AI Agents from Limited Interaction through Text-Tabular Modeling
di: Shapira, Eilam, et al.
Pubblicazione: (2026) -
Can LLMs Replace Economic Choice Prediction Labs? The Case of Language-based Persuasion Games
di: Shapira, Eilam, et al.
Pubblicazione: (2024)