Breaking the HISCO Barrier: Automatic Occupational Standardization with OccCANINE
Fuente:
arXiv
Guardado en:
| Autores principales: | Dahl, Christian Møller, Johansen, Torben, Vedel, Christian |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Neural Machine Translation for Malayalam Paraphrase Generation
por: Varghese, Christeena, et al.
Publicado: (2024)
por: Varghese, Christeena, et al.
Publicado: (2024)
Is Our Chatbot Telling Lies? Assessing Correctness of an LLM-based Dutch Support Chatbot
por: Lassche, Herman, et al.
Publicado: (2024)
por: Lassche, Herman, et al.
Publicado: (2024)
Predictive Simultaneous Interpretation: Harnessing Large Language Models for Democratizing Real-Time Multilingual Communication
por: Iida, Kurando, et al.
Publicado: (2024)
por: Iida, Kurando, et al.
Publicado: (2024)
UA-Legal-Bench: A Benchmark for Evaluating Large Language Models on Ukrainian Legal Reasoning
por: Ovcharov, Volodymyr
Publicado: (2026)
por: Ovcharov, Volodymyr
Publicado: (2026)
On Explaining with Attention Matrices
por: Naim, Omar, et al.
Publicado: (2024)
por: Naim, Omar, et al.
Publicado: (2024)
Inflation Attitudes of Large Language Models
por: Anesti, Nikoleta, et al.
Publicado: (2025)
por: Anesti, Nikoleta, et al.
Publicado: (2025)
Semantic Decomposition and Selective Context Filtering -- Text Processing Techniques for Context-Aware NLP-Based Systems
por: Villardar, Karl John
Publicado: (2025)
por: Villardar, Karl John
Publicado: (2025)
Classification of descriptions and summary using multiple passes of statistical and natural language toolkits
por: Banthia, Saumya, et al.
Publicado: (2020)
por: Banthia, Saumya, et al.
Publicado: (2020)
A new mapping of technological interdependence
por: Colladon, A. Fronzetti, et al.
Publicado: (2023)
por: Colladon, A. Fronzetti, et al.
Publicado: (2023)
LLMs and the Human Condition
por: Wallis, Peter
Publicado: (2024)
por: Wallis, Peter
Publicado: (2024)
Syntactic Blind Spots: How Misalignment Leads to LLMs Mathematical Errors
por: Williamson, Dane, et al.
Publicado: (2025)
por: Williamson, Dane, et al.
Publicado: (2025)
A Big Data Approach to Understand Sub-national Determinants of FDI in Africa
por: Colladon, A. Fronzetti, et al.
Publicado: (2024)
por: Colladon, A. Fronzetti, et al.
Publicado: (2024)
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
por: Ashuach, Tomer, et al.
Publicado: (2025)
por: Ashuach, Tomer, et al.
Publicado: (2025)
DECO-MWE: building a linguistic resource of Korean multiword expressions for feature-based sentiment analysis
por: Han, Jaeho, et al.
Publicado: (2026)
por: Han, Jaeho, et al.
Publicado: (2026)
Choosing features for classifying multiword expressions
por: Laporte, Eric
Publicado: (2026)
por: Laporte, Eric
Publicado: (2026)
Concordance Comparison as a Means of Assembling Local Grammars
por: Pirovani, Juliana, et al.
Publicado: (2026)
por: Pirovani, Juliana, et al.
Publicado: (2026)
Pattern-and-root inflectional morphology: the Arabic broken plural
por: Neme, Alexis Amid, et al.
Publicado: (2026)
por: Neme, Alexis Amid, et al.
Publicado: (2026)
What's Missing in Vision-Language Models? Probing Their Struggles with Causal Order Reasoning
por: Weng, Zhaotian, et al.
Publicado: (2025)
por: Weng, Zhaotian, et al.
Publicado: (2025)
Classification of non-analyzable word types in web documents to implement an effective Korean e-learning system
por: Park, Sang-Taek, et al.
Publicado: (2026)
por: Park, Sang-Taek, et al.
Publicado: (2026)
Where do aspectual variants of light verb constructions belong?
por: Fotopoulou, Aggeliki, et al.
Publicado: (2026)
por: Fotopoulou, Aggeliki, et al.
Publicado: (2026)
Conversion of Lexicon-Grammar tables to LMF. Application to French
por: Laporte, Eric, et al.
Publicado: (2026)
por: Laporte, Eric, et al.
Publicado: (2026)
Synthetic Voice Data for Automatic Speech Recognition in African Languages
por: DeRenzi, Brian, et al.
Publicado: (2025)
por: DeRenzi, Brian, et al.
Publicado: (2025)
TwinVoice: A Multi-dimensional Benchmark Towards Digital Twins via LLM Persona Simulation
por: Du, Bangde, et al.
Publicado: (2025)
por: Du, Bangde, et al.
Publicado: (2025)
ReFoRCE: A Text-to-SQL Agent with Self-Refinement, Consensus Enforcement, and Column Exploration
por: Deng, Minghang, et al.
Publicado: (2025)
por: Deng, Minghang, et al.
Publicado: (2025)
LegalGuardian: A Privacy-Preserving Framework for Secure Integration of Large Language Models in Legal Practice
por: Demir, M. Mikail, et al.
Publicado: (2025)
por: Demir, M. Mikail, et al.
Publicado: (2025)
ChemPro: A Progressive Chemistry Benchmark for Large Language Models
por: Baranwal, Aaditya, et al.
Publicado: (2026)
por: Baranwal, Aaditya, et al.
Publicado: (2026)
Towards Fundamental Language Models: Does Linguistic Competence Scale with Model Size?
por: Collado-Montañez, Jaime, et al.
Publicado: (2025)
por: Collado-Montañez, Jaime, et al.
Publicado: (2025)
SeLeRoSa: Sentence-Level Romanian Satire Detection Dataset
por: Smădu, Răzvan-Alexandru, et al.
Publicado: (2025)
por: Smădu, Răzvan-Alexandru, et al.
Publicado: (2025)
Generating training datasets for legal chatbots in Korean
por: Hwang, Changhoe, et al.
Publicado: (2026)
por: Hwang, Changhoe, et al.
Publicado: (2026)
SSP-based construction of evaluation-annotated data for fine-grained aspect-based sentiment analysis
por: Choi, Suwon, et al.
Publicado: (2026)
por: Choi, Suwon, et al.
Publicado: (2026)
French parsing enhanced with a word clustering method based on a syntactic lexicon
por: Sigogne, Anthony, et al.
Publicado: (2026)
por: Sigogne, Anthony, et al.
Publicado: (2026)
Graphemic Normalization of the Perso-Arabic Script
por: Doctor, Raiomond, et al.
Publicado: (2022)
por: Doctor, Raiomond, et al.
Publicado: (2022)
Beyond Arabic: Software for Perso-Arabic Script Manipulation
por: Gutkin, Alexander, et al.
Publicado: (2023)
por: Gutkin, Alexander, et al.
Publicado: (2023)
Aligning Language Models with Observational Data: Opportunities and Risks from a Causal Perspective
por: Loghmani, Erfan
Publicado: (2025)
por: Loghmani, Erfan
Publicado: (2025)
Quo Vadis ChatGPT? From Large Language Models to Large Knowledge Models
por: Venkatasubramanian, Venkat, et al.
Publicado: (2024)
por: Venkatasubramanian, Venkat, et al.
Publicado: (2024)
QuickSilver -- Speeding up LLM Inference through Dynamic Token Halting, KV Skipping, Contextual Token Fusion, and Adaptive Matryoshka Quantization
por: Khanna, Danush, et al.
Publicado: (2025)
por: Khanna, Danush, et al.
Publicado: (2025)
OpenAI Cribbed Our Tax Example, But Can GPT-4 Really Do Tax?
por: Blair-Stanek, Andrew, et al.
Publicado: (2023)
por: Blair-Stanek, Andrew, et al.
Publicado: (2023)
MATH-PT: A Math Reasoning Benchmark for European and Brazilian Portuguese
por: Teixeira, Tiago, et al.
Publicado: (2026)
por: Teixeira, Tiago, et al.
Publicado: (2026)
Semantic Commit: Helping Users Update Intent Specifications for AI Memory at Scale
por: Vaithilingam, Priyan, et al.
Publicado: (2025)
por: Vaithilingam, Priyan, et al.
Publicado: (2025)
Synergy of Large Language Model and Model Driven Engineering for Automated Development of Centralized Vehicular Systems
por: Petrovic, Nenad, et al.
Publicado: (2024)
por: Petrovic, Nenad, et al.
Publicado: (2024)
Ejemplares similares
-
Neural Machine Translation for Malayalam Paraphrase Generation
por: Varghese, Christeena, et al.
Publicado: (2024) -
Is Our Chatbot Telling Lies? Assessing Correctness of an LLM-based Dutch Support Chatbot
por: Lassche, Herman, et al.
Publicado: (2024) -
Predictive Simultaneous Interpretation: Harnessing Large Language Models for Democratizing Real-Time Multilingual Communication
por: Iida, Kurando, et al.
Publicado: (2024) -
UA-Legal-Bench: A Benchmark for Evaluating Large Language Models on Ukrainian Legal Reasoning
por: Ovcharov, Volodymyr
Publicado: (2026) -
On Explaining with Attention Matrices
por: Naim, Omar, et al.
Publicado: (2024)