CAST: Corpus-Aware Self-similarity Enhanced Topic modelling
Fuente:
arXiv
Saved in:
| Main Authors: | Ma, Yanan, Xiao, Chenghao, Yuan, Chenhan, van der Veer, Sabine N, Hassan, Lamiece, Lin, Chenghua, Nenadic, Goran |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Language Confusion Gate: Language-Aware Decoding Through Model Self-Distillation
by: Zhang, Collin, et al.
Published: (2025)
by: Zhang, Collin, et al.
Published: (2025)
Performing “Professionalism” in Grassroots Refugee Support: How Logics of Capital Enable Anti‐Migrant Hostility
by: Lieke van der Veer
Published: (2026)
by: Lieke van der Veer
Published: (2026)
The Achilles' Heel of Angular Margins: A Chebyshev Polynomial Fix for Speaker Verification
by: Wang, Yang, et al.
Published: (2026)
by: Wang, Yang, et al.
Published: (2026)
On the Rigour of Scientific Writing: Criteria, Analysis, and Insights
by: James, Joseph, et al.
Published: (2024)
by: James, Joseph, et al.
Published: (2024)
INSIGHTBUDDY-AI: Medication Extraction and Entity Linking using Large Language Models and Ensemble Learning
by: Romero, Pablo, et al.
Published: (2024)
by: Romero, Pablo, et al.
Published: (2024)
AutoLLM-CARD: Towards a Description and Landscape of Large Language Models
by: Tian, Shengwei, et al.
Published: (2024)
by: Tian, Shengwei, et al.
Published: (2024)
Clinical and cost‐effectiveness of telemedicine among patients with type 2 diabetes in primary care: A systematic review and meta‐analysis
by: Nawwarah Alfarwan, et al.
Published: (2024)
by: Nawwarah Alfarwan, et al.
Published: (2024)
Extract-and-Abstract: Unifying Extractive and Abstractive Summarization within Single Encoder-Decoder Framework
by: Wu, Yuping, et al.
Published: (2024)
by: Wu, Yuping, et al.
Published: (2024)
Investigating a Benchmark for Training-set free Evaluation of Linguistic Capabilities in Machine Reading Comprehension
by: Schlegel, Viktor, et al.
Published: (2024)
by: Schlegel, Viktor, et al.
Published: (2024)
Train & Constrain: Phonologically Informed Tongue-Twister Generation from Topics and Paraphrases
by: Loakman, Tyler, et al.
Published: (2024)
by: Loakman, Tyler, et al.
Published: (2024)
Effective Distillation of Table-based Reasoning Ability from LLMs
by: Yang, Bohao, et al.
Published: (2023)
by: Yang, Bohao, et al.
Published: (2023)
CAST: Non-Privileged Clipped Asymmetric Self-Teaching with Advantage Flipping for GRPO
by: Li, Yang, et al.
Published: (2026)
by: Li, Yang, et al.
Published: (2026)
Alexander Luria and Jean Piaget: Partial Reconstruction of Their Cooperation
by: Marc Ratcliff, et al.
Published: (2025)
by: Marc Ratcliff, et al.
Published: (2025)
De-identification of clinical free text using natural language processing: A systematic review of current approaches
by: Kovačević, Aleksandar, et al.
Published: (2023)
by: Kovačević, Aleksandar, et al.
Published: (2023)
Exploration of Masked and Causal Language Modelling for Text Generation
by: Micheletti, Nicolo, et al.
Published: (2024)
by: Micheletti, Nicolo, et al.
Published: (2024)
A Comparative Study on Automatic Coding of Medical Letters with Explainability
by: Glen, Jamie, et al.
Published: (2024)
by: Glen, Jamie, et al.
Published: (2024)
CAST: Clustering Self-Attention using Surrogate Tokens for Efficient Transformers
by: van Engelenhoven, Adjorn, et al.
Published: (2024)
by: van Engelenhoven, Adjorn, et al.
Published: (2024)
CAST: Cluster-Aware Self-Training for Tabular Data via Reliable Confidence
by: Kim, Minwook, et al.
Published: (2023)
by: Kim, Minwook, et al.
Published: (2023)
Owen-based Semantics and Hierarchy-Aware Explanation (O-Shap)
by: Zhou, Xiangyu, et al.
Published: (2026)
by: Zhou, Xiangyu, et al.
Published: (2026)
Comparing Apples to Oranges: A Dataset & Analysis of LLM Humour Understanding from Traditional Puns to Topical Jokes
by: Loakman, Tyler, et al.
Published: (2025)
by: Loakman, Tyler, et al.
Published: (2025)
Adversarial Defence without Adversarial Defence: Enhancing Language Model Robustness via Instance-level Principal Component Removal
by: Wang, Yang, et al.
Published: (2025)
by: Wang, Yang, et al.
Published: (2025)
Scalable and Reliable State-Aware Inference of High-Impact N-k Contingencies
by: Mai, Lihao, et al.
Published: (2026)
by: Mai, Lihao, et al.
Published: (2026)
RIGOURATE: Quantifying Scientific Exaggeration with Evidence-Aligned Claim Evaluation
by: James, Joseph, et al.
Published: (2026)
by: James, Joseph, et al.
Published: (2026)
Arg-LLaDA: Argument Summarization via Large Language Diffusion Models and Sufficiency-Aware Refinement
by: Li, Hao, et al.
Published: (2025)
by: Li, Hao, et al.
Published: (2025)
Computational techniques enabling the perception of virtual images exclusive to the retinal afterimage
by: de Jong, Staas, et al.
Published: (2025)
by: de Jong, Staas, et al.
Published: (2025)
V-CAST: Video Curvature-Aware Spatio-Temporal Pruning for Efficient Video Large Language Models
by: Lin, Xinying, et al.
Published: (2026)
by: Lin, Xinying, et al.
Published: (2026)
DeIDClinic: A Risk-Aware Pseudonymization Framework for Clinical Text De-identification and Re-identification Risk Assessment
by: Paul, Angel, et al.
Published: (2024)
by: Paul, Angel, et al.
Published: (2024)
Will Large Language Models Transform Clinical Prediction?
by: Yildiz, Yusuf, et al.
Published: (2025)
by: Yildiz, Yusuf, et al.
Published: (2025)
CAST: Collapse-Aware multi-Scale Topology Fusion for Multimodal Coreset Selection
by: Zhao, Boran, et al.
Published: (2026)
by: Zhao, Boran, et al.
Published: (2026)
Beyond One-Size-Fits-All: Inversion Learning for Highly Effective NLG Evaluation Prompts
by: Hong, Hanhua, et al.
Published: (2025)
by: Hong, Hanhua, et al.
Published: (2025)
Craftworkers in Nineteenth Century Scotland
by: Nenadic, Stana
Published: (2023)
by: Nenadic, Stana
Published: (2023)
Party‐Political Contestation of European Trade Policy. An Analysis of Roll Call Votes in the European Parliament
by: Simon Otjes, et al.
Published: (2025)
by: Simon Otjes, et al.
Published: (2025)
MTUncertainty: Assessing the Need for Post-editing of Machine Translation Outputs by Fine-tuning OpenAI LLMs
by: Gladkoff, Serge, et al.
Published: (2023)
by: Gladkoff, Serge, et al.
Published: (2023)
CANTONMT: Investigating Back-Translation and Model-Switch Mechanisms for Cantonese-English Neural Machine Translation
by: Hong, Kung Yin, et al.
Published: (2024)
by: Hong, Kung Yin, et al.
Published: (2024)
CantonMT: Cantonese to English NMT Platform with Fine-Tuned Models Using Synthetic Back-Translation Data
by: Hong, Kung Yin, et al.
Published: (2024)
by: Hong, Kung Yin, et al.
Published: (2024)
MaLei at the PLABA Track of TREC 2024: RoBERTa for Term Replacement -- LLaMA3.1 and GPT-4o for Complete Abstract Adaptation
by: Ling, Zhidong, et al.
Published: (2024)
by: Ling, Zhidong, et al.
Published: (2024)
HealthcareNLP: where are we and what is next?
by: Han, Lifeng, et al.
Published: (2025)
by: Han, Lifeng, et al.
Published: (2025)
Generating Synthetic Free-text Medical Records with Low Re-identification Risk using Masked Language Modeling
by: Belkadi, Samuel, et al.
Published: (2024)
by: Belkadi, Samuel, et al.
Published: (2024)
Property-Preserving Hashing for $\ell_1$-Distance Predicates: Applications to Countering Adversarial Input Attacks
by: Asghar, Hassan, et al.
Published: (2025)
by: Asghar, Hassan, et al.
Published: (2025)
New Grounds for Ontic Trust: Information Objects and LIS
by: Van der Veer Martens, Betsy
Published: (2017)
by: Van der Veer Martens, Betsy
Published: (2017)
Similar Items
-
Language Confusion Gate: Language-Aware Decoding Through Model Self-Distillation
by: Zhang, Collin, et al.
Published: (2025) -
Performing “Professionalism” in Grassroots Refugee Support: How Logics of Capital Enable Anti‐Migrant Hostility
by: Lieke van der Veer
Published: (2026) -
The Achilles' Heel of Angular Margins: A Chebyshev Polynomial Fix for Speaker Verification
by: Wang, Yang, et al.
Published: (2026) -
On the Rigour of Scientific Writing: Criteria, Analysis, and Insights
by: James, Joseph, et al.
Published: (2024) -
INSIGHTBUDDY-AI: Medication Extraction and Entity Linking using Large Language Models and Ensemble Learning
by: Romero, Pablo, et al.
Published: (2024)