FlexDoc: Parameterized Sampling for Diverse Multilingual Synthetic Documents for Training Document Understanding Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Dua, Karan, Patel, Hitesh Laxmichand, Mittal, Puneet, Gupta, Ranjeet, Agarwal, Amit, Pabolu, Praneet, Panda, Srikant, Meghwani, Hansa, Horwood, Graham, Shah, Fahad |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PCRI: Measuring Context Robustness in Multimodal Models for Enterprise Applications
von: Patel, Hitesh Laxmichand, et al.
Veröffentlicht: (2025)
von: Patel, Hitesh Laxmichand, et al.
Veröffentlicht: (2025)
RCI: A Score for Evaluating Global and Local Reasoning in Multimodal Benchmarks
von: Agarwal, Amit, et al.
Veröffentlicht: (2025)
von: Agarwal, Amit, et al.
Veröffentlicht: (2025)
Hard Negative Mining for Domain-Specific Retrieval in Enterprise Systems
von: Meghwani, Hansa, et al.
Veröffentlicht: (2025)
von: Meghwani, Hansa, et al.
Veröffentlicht: (2025)
MVTamperBench: Evaluating Robustness of Vision-Language Models
von: Agarwal, Amit, et al.
Veröffentlicht: (2024)
von: Agarwal, Amit, et al.
Veröffentlicht: (2024)
Aligning LLMs for Multilingual Consistency in Enterprise Applications
von: Agarwal, Amit, et al.
Veröffentlicht: (2025)
von: Agarwal, Amit, et al.
Veröffentlicht: (2025)
Enhancing Document AI Data Generation Through Graph-Based Synthetic Layouts
von: Agarwal, Amit, et al.
Veröffentlicht: (2024)
von: Agarwal, Amit, et al.
Veröffentlicht: (2024)
Who's Asking? Investigating Bias Through the Lens of Disability Framed Queries in LLMs
von: Hari, Vishnu, et al.
Veröffentlicht: (2025)
von: Hari, Vishnu, et al.
Veröffentlicht: (2025)
SpeechWeave: Diverse Multilingual Synthetic Text & Audio Data Generation Pipeline for Training Text to Speech Models
von: Dua, Karan, et al.
Veröffentlicht: (2025)
von: Dua, Karan, et al.
Veröffentlicht: (2025)
FS-DAG: Few Shot Domain Adapting Graph Networks for Visually Rich Document Understanding
von: Agarwal, Amit, et al.
Veröffentlicht: (2025)
von: Agarwal, Amit, et al.
Veröffentlicht: (2025)
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
von: Raoufi, Behnam, et al.
Veröffentlicht: (2025)
von: Raoufi, Behnam, et al.
Veröffentlicht: (2025)
CCVA-FL: Cross-Client Variations Adaptive Federated Learning for Medical Imaging
von: Gupta, Sunny, et al.
Veröffentlicht: (2024)
von: Gupta, Sunny, et al.
Veröffentlicht: (2024)
Taming the Tail: Leveraging Asymmetric Loss and Pade Approximation to Overcome Medical Image Long-Tailed Class Imbalance
von: Kashyap, Pankhi, et al.
Veröffentlicht: (2024)
von: Kashyap, Pankhi, et al.
Veröffentlicht: (2024)
FLD+: Data-efficient Evaluation Metric for Generative Models
von: Jeevan, Pranav, et al.
Veröffentlicht: (2024)
von: Jeevan, Pranav, et al.
Veröffentlicht: (2024)
WaveMixSR-V2: Enhancing Super-resolution with Higher Efficiency
von: Jeevan, Pranav, et al.
Veröffentlicht: (2024)
von: Jeevan, Pranav, et al.
Veröffentlicht: (2024)
Normalizing Flow-Based Metric for Image Generation
von: Jeevan, Pranav, et al.
Veröffentlicht: (2024)
von: Jeevan, Pranav, et al.
Veröffentlicht: (2024)
WaveMix: A Resource-efficient Neural Network for Image Analysis
von: Jeevan, Pranav, et al.
Veröffentlicht: (2022)
von: Jeevan, Pranav, et al.
Veröffentlicht: (2022)
Which Backbone to Use: A Resource-efficient Domain Specific Comparison for Computer Vision
von: Jeevan, Pranav, et al.
Veröffentlicht: (2024)
von: Jeevan, Pranav, et al.
Veröffentlicht: (2024)
Scaling Laws for State Dynamics in Large Language Models
von: Li, Jacob X, et al.
Veröffentlicht: (2025)
von: Li, Jacob X, et al.
Veröffentlicht: (2025)
FEDTAIL: Federated Long-Tailed Domain Generalization with Sharpness-Guided Gradient Matching
von: Gupta, Sunny, et al.
Veröffentlicht: (2025)
von: Gupta, Sunny, et al.
Veröffentlicht: (2025)
DeepC4: Deep Conditional Census-Constrained Clustering for Large-scale Multitask Spatial Disaggregation of Urban Morphology
von: Dimasaka, Joshua, et al.
Veröffentlicht: (2025)
von: Dimasaka, Joshua, et al.
Veröffentlicht: (2025)
Evaluation Metric for Quality Control and Generative Models in Histopathology Images
von: Jeevan, Pranav, et al.
Veröffentlicht: (2024)
von: Jeevan, Pranav, et al.
Veröffentlicht: (2024)
FedStein: Enhancing Multi-Domain Federated Learning Through James-Stein Estimator
von: Gupta, Sunny, et al.
Veröffentlicht: (2024)
von: Gupta, Sunny, et al.
Veröffentlicht: (2024)
UniVarFL: Uniformity and Variance Regularized Federated Learning for Heterogeneous Data
von: Gupta, Sunny, et al.
Veröffentlicht: (2025)
von: Gupta, Sunny, et al.
Veröffentlicht: (2025)
FedAlign: Federated Domain Generalization with Cross-Client Feature Alignment
von: Gupta, Sunny, et al.
Veröffentlicht: (2025)
von: Gupta, Sunny, et al.
Veröffentlicht: (2025)
Trapped in texture bias? A large scale comparison of deep instance segmentation
von: Theodoridis, Johannes, et al.
Veröffentlicht: (2024)
von: Theodoridis, Johannes, et al.
Veröffentlicht: (2024)
Parameterizing Dataset Distillation via Gaussian Splatting
von: Jiang, Chenyang, et al.
Veröffentlicht: (2025)
von: Jiang, Chenyang, et al.
Veröffentlicht: (2025)
propella-1: Multi-Property Document Annotation for LLM Data Curation at Scale
von: Idahl, Maximilian, et al.
Veröffentlicht: (2026)
von: Idahl, Maximilian, et al.
Veröffentlicht: (2026)
LongSumEval: Question-Answering Based Evaluation and Feedback-Driven Refinement for Long Document Summarization
von: Nguyen, Huyen, et al.
Veröffentlicht: (2026)
von: Nguyen, Huyen, et al.
Veröffentlicht: (2026)
Few-Shot Learning of a Graph-Based Neural Network Model Without Backpropagation
von: Lapin, Mykyta, et al.
Veröffentlicht: (2025)
von: Lapin, Mykyta, et al.
Veröffentlicht: (2025)
Sparse Concept Bottleneck Models: Gumbel Tricks in Contrastive Learning
von: Semenov, Andrei, et al.
Veröffentlicht: (2024)
von: Semenov, Andrei, et al.
Veröffentlicht: (2024)
GLoT: A Novel Gated-Logarithmic Transformer for Efficient Sign Language Translation
von: Shahin, Nada, et al.
Veröffentlicht: (2025)
von: Shahin, Nada, et al.
Veröffentlicht: (2025)
Explainable Classifier for Malignant Lymphoma Subtyping via Cell Graph and Image Fusion
von: Nishiyama, Daiki, et al.
Veröffentlicht: (2025)
von: Nishiyama, Daiki, et al.
Veröffentlicht: (2025)
A Roadmap for Multilingual, Multimodal Domain Independent Deception Detection
von: Boumber, Dainis, et al.
Veröffentlicht: (2024)
von: Boumber, Dainis, et al.
Veröffentlicht: (2024)
When in Doubt, Think Slow: Iterative Reasoning with Latent Imagination
von: Benfeghoul, Martin, et al.
Veröffentlicht: (2024)
von: Benfeghoul, Martin, et al.
Veröffentlicht: (2024)
Context-dependent Causality (the Non-Nonotonic Case)
von: Billfeld, Nir, et al.
Veröffentlicht: (2024)
von: Billfeld, Nir, et al.
Veröffentlicht: (2024)
Interpreting Structured Perturbations in Image Protection Methods for Diffusion Models
von: Martin, Michael R., et al.
Veröffentlicht: (2025)
von: Martin, Michael R., et al.
Veröffentlicht: (2025)
IMUVIE: Pickup Timeline Action Localization via Motion Movies
von: Clapham, John, et al.
Veröffentlicht: (2024)
von: Clapham, John, et al.
Veröffentlicht: (2024)
Supervised Embedded Methods for Hyperspectral Band Selection
von: Zimmer, Yaniv, et al.
Veröffentlicht: (2024)
von: Zimmer, Yaniv, et al.
Veröffentlicht: (2024)
CC-SGG: Corner Case Scenario Generation using Learned Scene Graphs
von: Drayson, George, et al.
Veröffentlicht: (2023)
von: Drayson, George, et al.
Veröffentlicht: (2023)
FSFM: A Biologically-Inspired Framework for Selective Forgetting of Agent Memory
von: Gu, Yingjie, et al.
Veröffentlicht: (2026)
von: Gu, Yingjie, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
PCRI: Measuring Context Robustness in Multimodal Models for Enterprise Applications
von: Patel, Hitesh Laxmichand, et al.
Veröffentlicht: (2025) -
RCI: A Score for Evaluating Global and Local Reasoning in Multimodal Benchmarks
von: Agarwal, Amit, et al.
Veröffentlicht: (2025) -
Hard Negative Mining for Domain-Specific Retrieval in Enterprise Systems
von: Meghwani, Hansa, et al.
Veröffentlicht: (2025) -
MVTamperBench: Evaluating Robustness of Vision-Language Models
von: Agarwal, Amit, et al.
Veröffentlicht: (2024) -
Aligning LLMs for Multilingual Consistency in Enterprise Applications
von: Agarwal, Amit, et al.
Veröffentlicht: (2025)