ATLAS: Benchmarking and Adapting LLMs for Global Trade via Harmonized Tariff Code Classification
Fuente:
arXiv
Salvato in:
| Autori principali: | Yuvraj, Pritish, Devarakonda, Siva |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Benchmarking Harmonized Tariff Schedule Classification Models
di: Judy, Bryce
Pubblicazione: (2024)
di: Judy, Bryce
Pubblicazione: (2024)
Benchmarking and Adapting On-Device LLMs for Clinical Decision Support
di: Munim, Alif, et al.
Pubblicazione: (2025)
di: Munim, Alif, et al.
Pubblicazione: (2025)
REGAL: A Registry-Driven Architecture for Deterministic Grounding of Agentic AI in Enterprise Telemetry
di: Agrawal, Yuvraj
Pubblicazione: (2026)
di: Agrawal, Yuvraj
Pubblicazione: (2026)
Layer Swapping for Zero-Shot Cross-Lingual Transfer in Large Language Models
di: Bandarkar, Lucas, et al.
Pubblicazione: (2024)
di: Bandarkar, Lucas, et al.
Pubblicazione: (2024)
A Deterministic Agentic Workflow for HS Tariff Classification: Multi-Dimensional Rule Reasoning with Interpretable Decisions
di: Zhang, Yu, et al.
Pubblicazione: (2026)
di: Zhang, Yu, et al.
Pubblicazione: (2026)
AdaptEval: A Benchmark for Evaluating Large Language Models on Code Snippet Adaptation
di: Zhang, Tanghaoran, et al.
Pubblicazione: (2026)
di: Zhang, Tanghaoran, et al.
Pubblicazione: (2026)
DistShap: Scalable GNN Explanations with Distributed Shapley Values
di: Akkas, Selahattin, et al.
Pubblicazione: (2025)
di: Akkas, Selahattin, et al.
Pubblicazione: (2025)
Classification-Based Automatic HDL Code Generation Using LLMs
di: Sun, Wenhao, et al.
Pubblicazione: (2024)
di: Sun, Wenhao, et al.
Pubblicazione: (2024)
Can We Make Code Green? Understanding Trade-Offs in LLMs vs. Human Code Optimizations
di: Rani, Pooja, et al.
Pubblicazione: (2025)
di: Rani, Pooja, et al.
Pubblicazione: (2025)
ResearchCodeBench: Benchmarking LLMs on Implementing Novel Machine Learning Research Code
di: Hua, Tianyu, et al.
Pubblicazione: (2025)
di: Hua, Tianyu, et al.
Pubblicazione: (2025)
Hallucination by Code Generation LLMs: Taxonomy, Benchmarks, Mitigation, and Challenges
di: Lee, Yunseo, et al.
Pubblicazione: (2025)
di: Lee, Yunseo, et al.
Pubblicazione: (2025)
Benchmarking Multimodal LLMs on Code Generation for Complex Interactive Webpages
di: Wu, Fan, et al.
Pubblicazione: (2026)
di: Wu, Fan, et al.
Pubblicazione: (2026)
PythonSaga: Redefining the Benchmark to Evaluate Code Generating LLMs
di: Yadav, Ankit, et al.
Pubblicazione: (2024)
di: Yadav, Ankit, et al.
Pubblicazione: (2024)
Adapting Multilingual LLMs to Low-Resource Languages with Knowledge Graphs via Adapters
di: Gurgurov, Daniil, et al.
Pubblicazione: (2024)
di: Gurgurov, Daniil, et al.
Pubblicazione: (2024)
RAIL in the Wild: Operationalizing Responsible AI Evaluation Using Anthropic's Value Dataset
di: Verma, Sumit, et al.
Pubblicazione: (2025)
di: Verma, Sumit, et al.
Pubblicazione: (2025)
Multimodal Approach for Harmonized System Code Prediction
di: Amel, Otmane, et al.
Pubblicazione: (2024)
di: Amel, Otmane, et al.
Pubblicazione: (2024)
Benchmarking LLMs for Fine-Grained Code Review with Enriched Context in Practice
di: Hu, Ruida, et al.
Pubblicazione: (2025)
di: Hu, Ruida, et al.
Pubblicazione: (2025)
Beyond Code Snippets: Benchmarking LLMs on Repository-Level Question Answering
di: Alebachew, Yoseph Berhanu, et al.
Pubblicazione: (2026)
di: Alebachew, Yoseph Berhanu, et al.
Pubblicazione: (2026)
Adapting LLMs to Time Series Forecasting via Temporal Heterogeneity Modeling and Semantic Alignment
di: Sun, Yanru, et al.
Pubblicazione: (2025)
di: Sun, Yanru, et al.
Pubblicazione: (2025)
Adapting LLMs for Minimal-edit Grammatical Error Correction
di: Staruch, Ryszard, et al.
Pubblicazione: (2025)
di: Staruch, Ryszard, et al.
Pubblicazione: (2025)
Are LLMs Ready for TOON? Benchmarking Structural Correctness-Sustainability Trade-offs in Novel Structured Output Formats
di: Masciari, Elio, et al.
Pubblicazione: (2026)
di: Masciari, Elio, et al.
Pubblicazione: (2026)
EduAdapt: A Question Answer Benchmark Dataset for Evaluating Grade-Level Adaptability in LLMs
di: Naeem, Numaan, et al.
Pubblicazione: (2025)
di: Naeem, Numaan, et al.
Pubblicazione: (2025)
Rectifier: Code Translation with Corrector via LLMs
di: Yin, Xin, et al.
Pubblicazione: (2024)
di: Yin, Xin, et al.
Pubblicazione: (2024)
HardSecBench: Benchmarking the Security Awareness of LLMs for Hardware Code Generation
di: Chen, Qirui, et al.
Pubblicazione: (2026)
di: Chen, Qirui, et al.
Pubblicazione: (2026)
Validate Your Authority: Benchmarking LLMs on Multi-Label Precedent Treatment Classification
di: Demir, M. Mikail, et al.
Pubblicazione: (2026)
di: Demir, M. Mikail, et al.
Pubblicazione: (2026)
Smaller = Weaker? Benchmarking Robustness of Quantized LLMs in Code Generation
di: Fang, Sen, et al.
Pubblicazione: (2025)
di: Fang, Sen, et al.
Pubblicazione: (2025)
Automating Code Adaptation for MLOps -- A Benchmarking Study on LLMs
di: Patel, Harsh, et al.
Pubblicazione: (2024)
di: Patel, Harsh, et al.
Pubblicazione: (2024)
Drawing Pandas: A Benchmark for LLMs in Generating Plotting Code
di: Galimzyanov, Timur, et al.
Pubblicazione: (2024)
di: Galimzyanov, Timur, et al.
Pubblicazione: (2024)
SensorBench: Benchmarking LLMs in Coding-Based Sensor Processing
di: Quan, Pengrui, et al.
Pubblicazione: (2024)
di: Quan, Pengrui, et al.
Pubblicazione: (2024)
Unsupervised Learning of Harmonic Analysis Based on Neural HSMM with Code Quality Templates
di: Uehara, Yui
Pubblicazione: (2024)
di: Uehara, Yui
Pubblicazione: (2024)
When Developer Aid Becomes Security Debt: A Systematic Analysis of Insecure Behaviors in LLM Coding Agents
di: Kozak, Matous, et al.
Pubblicazione: (2025)
di: Kozak, Matous, et al.
Pubblicazione: (2025)
ATLAS: Adaptive Trading with LLM AgentS Through Dynamic Prompt Optimization and Multi-Agent Coordination
di: Papadakis, Charidimos, et al.
Pubblicazione: (2025)
di: Papadakis, Charidimos, et al.
Pubblicazione: (2025)
ACT: Bridging the Gap in Code Translation through Synthetic Data Generation & Adaptive Training
di: Saxena, Shreya, et al.
Pubblicazione: (2025)
di: Saxena, Shreya, et al.
Pubblicazione: (2025)
Harmonic LLMs are Trustworthy
di: Kersting, Nicholas S., et al.
Pubblicazione: (2024)
di: Kersting, Nicholas S., et al.
Pubblicazione: (2024)
EnvGen: Generating and Adapting Environments via LLMs for Training Embodied Agents
di: Zala, Abhay, et al.
Pubblicazione: (2024)
di: Zala, Abhay, et al.
Pubblicazione: (2024)
Do LLMs Really Adapt to Domains? An Ontology Learning Perspective
di: Mai, Huu Tan, et al.
Pubblicazione: (2024)
di: Mai, Huu Tan, et al.
Pubblicazione: (2024)
Benchmark Health Index: A Systematic Framework for Benchmarking the Benchmarks of LLMs
di: Zhu, Longyuan, et al.
Pubblicazione: (2026)
di: Zhu, Longyuan, et al.
Pubblicazione: (2026)
CoRe: Benchmarking LLMs Code Reasoning Capabilities through Static Analysis Tasks
di: Xie, Danning, et al.
Pubblicazione: (2025)
di: Xie, Danning, et al.
Pubblicazione: (2025)
Benchmarking Overton Pluralism in LLMs
di: Poole-Dayan, Elinor, et al.
Pubblicazione: (2025)
di: Poole-Dayan, Elinor, et al.
Pubblicazione: (2025)
SpecRover: Code Intent Extraction via LLMs
di: Ruan, Haifeng, et al.
Pubblicazione: (2024)
di: Ruan, Haifeng, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Benchmarking Harmonized Tariff Schedule Classification Models
di: Judy, Bryce
Pubblicazione: (2024) -
Benchmarking and Adapting On-Device LLMs for Clinical Decision Support
di: Munim, Alif, et al.
Pubblicazione: (2025) -
REGAL: A Registry-Driven Architecture for Deterministic Grounding of Agentic AI in Enterprise Telemetry
di: Agrawal, Yuvraj
Pubblicazione: (2026) -
Layer Swapping for Zero-Shot Cross-Lingual Transfer in Large Language Models
di: Bandarkar, Lucas, et al.
Pubblicazione: (2024) -
A Deterministic Agentic Workflow for HS Tariff Classification: Multi-Dimensional Rule Reasoning with Interpretable Decisions
di: Zhang, Yu, et al.
Pubblicazione: (2026)