AI-assisted German Employment Contract Review: A Benchmark Dataset
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wardas, Oliver, Matthes, Florian |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LLMs for Legal Subsumption in German Employment Contracts
von: Wardas, Oliver, et al.
Veröffentlicht: (2025)
von: Wardas, Oliver, et al.
Veröffentlicht: (2025)
PaperAudit-Bench: Benchmarking Error Detection in Research Papers for Critical Automated Peer Review
von: Tu, Songjun, et al.
Veröffentlicht: (2026)
von: Tu, Songjun, et al.
Veröffentlicht: (2026)
Profiling German Text Simplification with Interpretable Model-Fingerprints
von: Klöser, Lars, et al.
Veröffentlicht: (2026)
von: Klöser, Lars, et al.
Veröffentlicht: (2026)
Omni-SafetyBench: A Benchmark for Safety Evaluation of Audio-Visual Large Language Models
von: Pan, Leyi, et al.
Veröffentlicht: (2025)
von: Pan, Leyi, et al.
Veröffentlicht: (2025)
Compression Method Matters: Benchmark-Dependent Output Dynamics in LLM Prompt Compression
von: Johnson, Warren
Veröffentlicht: (2026)
von: Johnson, Warren
Veröffentlicht: (2026)
IFMTBench: A Comprehensive Benchmark for Multilingual Translation Instruction Following
von: Sun, Mingrui, et al.
Veröffentlicht: (2026)
von: Sun, Mingrui, et al.
Veröffentlicht: (2026)
Co-NAML-LSTUR: A Combined Model with Attentive Multi-View Learning and Long- and Short-term User Representations for News Recommendation
von: Nguyen, Minh Hoang, et al.
Veröffentlicht: (2025)
von: Nguyen, Minh Hoang, et al.
Veröffentlicht: (2025)
A Survey on Natural Language Counterfactual Generation
von: Wang, Yongjie, et al.
Veröffentlicht: (2024)
von: Wang, Yongjie, et al.
Veröffentlicht: (2024)
A Survey of Text Watermarking in the Era of Large Language Models
von: Liu, Aiwei, et al.
Veröffentlicht: (2023)
von: Liu, Aiwei, et al.
Veröffentlicht: (2023)
A Stochastic Analysis of the Linguistic Provenance of English Place Names
von: Dalvean, Michael
Veröffentlicht: (2023)
von: Dalvean, Michael
Veröffentlicht: (2023)
New Skills or Sharper Primitives? A Probabilistic Perspective on the Emergence of Reasoning in RLVR
von: Wang, Zhilin, et al.
Veröffentlicht: (2026)
von: Wang, Zhilin, et al.
Veröffentlicht: (2026)
Prompt Compression in Production Task Orchestration: A Pre-Registered Randomized Trial
von: Johnson, Warren, et al.
Veröffentlicht: (2026)
von: Johnson, Warren, et al.
Veröffentlicht: (2026)
Relating Word Embedding Gender Biases to Gender Gaps: A Cross-Cultural Analysis
von: Friedman, Scott, et al.
Veröffentlicht: (2026)
von: Friedman, Scott, et al.
Veröffentlicht: (2026)
Large Language Models(LLMs) on Tabular Data: Prediction, Generation, and Understanding -- A Survey
von: Fang, Xi, et al.
Veröffentlicht: (2024)
von: Fang, Xi, et al.
Veröffentlicht: (2024)
Unveiling Attractor Cycles in Large Language Models: A Dynamical Systems View of Successive Paraphrasing
von: Wang, Zhilin, et al.
Veröffentlicht: (2025)
von: Wang, Zhilin, et al.
Veröffentlicht: (2025)
Trusted Uncertainty in Large Language Models: A Unified Framework for Confidence Calibration and Risk-Controlled Refusal
von: Oehri, Markus, et al.
Veröffentlicht: (2025)
von: Oehri, Markus, et al.
Veröffentlicht: (2025)
SentiCSE: A Sentiment-aware Contrastive Sentence Embedding Framework with Sentiment-guided Textual Similarity
von: Kim, Jaemin, et al.
Veröffentlicht: (2024)
von: Kim, Jaemin, et al.
Veröffentlicht: (2024)
ScoreRAG: A Retrieval-Augmented Generation Framework with Consistency-Relevance Scoring and Structured Summarization for News Generation
von: Lin, Pei-Yun, et al.
Veröffentlicht: (2025)
von: Lin, Pei-Yun, et al.
Veröffentlicht: (2025)
Examining Linguistic Shifts in Academic Writing Before and After the Launch of ChatGPT: A Study on Preprint Papers
von: Bao, Tong, et al.
Veröffentlicht: (2025)
von: Bao, Tong, et al.
Veröffentlicht: (2025)
Math Natural Language Inference: this should be easy!
von: de Paiva, Valeria, et al.
Veröffentlicht: (2025)
von: de Paiva, Valeria, et al.
Veröffentlicht: (2025)
Fast Quiet-STaR: Thinking Without Thought Tokens
von: Huang, Wei, et al.
Veröffentlicht: (2025)
von: Huang, Wei, et al.
Veröffentlicht: (2025)
d-TreeRPO: Towards More Reliable Policy Optimization for Diffusion Language Models
von: Pan, Leyi, et al.
Veröffentlicht: (2025)
von: Pan, Leyi, et al.
Veröffentlicht: (2025)
Unifying Uniform and Binary-coding Quantization for Accurate Compression of Large Language Models
von: Park, Seungcheol, et al.
Veröffentlicht: (2025)
von: Park, Seungcheol, et al.
Veröffentlicht: (2025)
When Retrieval Succeeds and Fails: Rethinking Retrieval-Augmented Generation for LLMs
von: Wang, Yongjie, et al.
Veröffentlicht: (2025)
von: Wang, Yongjie, et al.
Veröffentlicht: (2025)
GATE: Graph-based Adaptive Tool Evolution Across Diverse Tasks
von: Luo, Jianwen, et al.
Veröffentlicht: (2025)
von: Luo, Jianwen, et al.
Veröffentlicht: (2025)
Culturally-Nuanced Story Generation for Reasoning in Low-Resource Languages: The Case of Javanese and Sundanese
von: Pranida, Salsabila Zahirah, et al.
Veröffentlicht: (2025)
von: Pranida, Salsabila Zahirah, et al.
Veröffentlicht: (2025)
The Paradox of Poetic Intent in Back-Translation: Evaluating the Quality of Large Language Models in Chinese Translation
von: Weigang, Li, et al.
Veröffentlicht: (2025)
von: Weigang, Li, et al.
Veröffentlicht: (2025)
Can LLM Watermarks Robustly Prevent Unauthorized Knowledge Distillation?
von: Pan, Leyi, et al.
Veröffentlicht: (2025)
von: Pan, Leyi, et al.
Veröffentlicht: (2025)
Charting a Decade of Computational Linguistics in Italy: The CLiC-it Corpus
von: Alzetta, Chiara, et al.
Veröffentlicht: (2025)
von: Alzetta, Chiara, et al.
Veröffentlicht: (2025)
Exploring State Tracking Capabilities of Large Language Models
von: Rezaee, Kiamehr, et al.
Veröffentlicht: (2025)
von: Rezaee, Kiamehr, et al.
Veröffentlicht: (2025)
Quantifying Genuine Awareness in Hallucination Prediction Beyond Question-Side Shortcuts
von: Seo, Yeongbin, et al.
Veröffentlicht: (2025)
von: Seo, Yeongbin, et al.
Veröffentlicht: (2025)
NERCat: Fine-Tuning for Enhanced Named Entity Recognition in Catalan
von: Ferreres, Guillem Cadevall, et al.
Veröffentlicht: (2025)
von: Ferreres, Guillem Cadevall, et al.
Veröffentlicht: (2025)
Accurate Sublayer Pruning for Large Language Models by Exploiting Latency and Tunability Information
von: Park, Seungcheol, et al.
Veröffentlicht: (2025)
von: Park, Seungcheol, et al.
Veröffentlicht: (2025)
"As Eastern Powers, I will veto." : An Investigation of Nation-level Bias of Large Language Models in International Relations
von: Choi, Jonghyeon, et al.
Veröffentlicht: (2025)
von: Choi, Jonghyeon, et al.
Veröffentlicht: (2025)
Testing the assumptions about the geometry of sentence embedding spaces: the cosine measure need not apply
von: Nastase, Vivi, et al.
Veröffentlicht: (2025)
von: Nastase, Vivi, et al.
Veröffentlicht: (2025)
Nested Named Entity Recognition as Single-Pass Sequence Labeling
von: Muñoz-Ortiz, Alberto, et al.
Veröffentlicht: (2025)
von: Muñoz-Ortiz, Alberto, et al.
Veröffentlicht: (2025)
NurValues: Real-World Nursing Values Evaluation for Large Language Models in Clinical Context
von: Yao, Ben, et al.
Veröffentlicht: (2025)
von: Yao, Ben, et al.
Veröffentlicht: (2025)
WeDLM: Reconciling Diffusion Language Models with Standard Causal Attention for Fast Inference
von: Liu, Aiwei, et al.
Veröffentlicht: (2025)
von: Liu, Aiwei, et al.
Veröffentlicht: (2025)
Prior-based Noisy Text Data Filtering: Fast and Strong Alternative For Perplexity
von: Seo, Yeongbin, et al.
Veröffentlicht: (2025)
von: Seo, Yeongbin, et al.
Veröffentlicht: (2025)
Disambiguation of Emotion Annotations by Contextualizing Events in Plausible Narratives
von: Schäfer, Johannes, et al.
Veröffentlicht: (2025)
von: Schäfer, Johannes, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
LLMs for Legal Subsumption in German Employment Contracts
von: Wardas, Oliver, et al.
Veröffentlicht: (2025) -
PaperAudit-Bench: Benchmarking Error Detection in Research Papers for Critical Automated Peer Review
von: Tu, Songjun, et al.
Veröffentlicht: (2026) -
Profiling German Text Simplification with Interpretable Model-Fingerprints
von: Klöser, Lars, et al.
Veröffentlicht: (2026) -
Omni-SafetyBench: A Benchmark for Safety Evaluation of Audio-Visual Large Language Models
von: Pan, Leyi, et al.
Veröffentlicht: (2025) -
Compression Method Matters: Benchmark-Dependent Output Dynamics in LLM Prompt Compression
von: Johnson, Warren
Veröffentlicht: (2026)