Corrections of Zipf's and Heaps' Laws Derived from Hapax Rate Models
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Dębowski, Łukasz |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Word Embedding Techniques for Classification of Star Ratings
von: Abdelmotaleb, Hesham, et al.
Veröffentlicht: (2025)
von: Abdelmotaleb, Hesham, et al.
Veröffentlicht: (2025)
From Noisy News Sentiment Scores to Interpretable Temporal Dynamics: A Bayesian State-Space Model
von: Casals, Ian Carbó
Veröffentlicht: (2026)
von: Casals, Ian Carbó
Veröffentlicht: (2026)
A Structural Text-Based Scaling Model for Analyzing Political Discourse
von: Vávra, Jan, et al.
Veröffentlicht: (2024)
von: Vávra, Jan, et al.
Veröffentlicht: (2024)
Claim Automation using Large Language Model
von: Mo, Zhengda, et al.
Veröffentlicht: (2026)
von: Mo, Zhengda, et al.
Veröffentlicht: (2026)
Pre-trained Models Perform the Best When Token Distributions Follow Zipf's Law
von: He, Yanjin, et al.
Veröffentlicht: (2025)
von: He, Yanjin, et al.
Veröffentlicht: (2025)
A methodological analysis of prompt perturbations and their effect on attack success rates
von: Machado, Tiago, et al.
Veröffentlicht: (2025)
von: Machado, Tiago, et al.
Veröffentlicht: (2025)
Multidimensional Analysis of Specific Language Impairment Using Unsupervised Learning Through PCA and Clustering
von: Selvanayagam, Niruthiha
Veröffentlicht: (2025)
von: Selvanayagam, Niruthiha
Veröffentlicht: (2025)
Conformal Prediction Sets for Next-Token Prediction in Large Language Models: Balancing Coverage Guarantees with Set Efficiency
von: Kotla, Yoshith Roy, et al.
Veröffentlicht: (2025)
von: Kotla, Yoshith Roy, et al.
Veröffentlicht: (2025)
Dynamical Survival Analysis for Modeling Hazard Functions with Nonlinear Systems
von: Liyanage, Dananjani, et al.
Veröffentlicht: (2026)
von: Liyanage, Dananjani, et al.
Veröffentlicht: (2026)
Automated Quality Assessment for LLM-Based Complex Qualitative Coding: A Confidence-Diversity Framework
von: Zhao, Zhilong, et al.
Veröffentlicht: (2025)
von: Zhao, Zhilong, et al.
Veröffentlicht: (2025)
An Information-Theoretic Approach for Detecting Edits in AI-Generated Text
von: Kashtan, Idan, et al.
Veröffentlicht: (2023)
von: Kashtan, Idan, et al.
Veröffentlicht: (2023)
Towards Lighter and Robust Evaluation for Retrieval Augmented Generation
von: Ispas, Alex-Razvan, et al.
Veröffentlicht: (2025)
von: Ispas, Alex-Razvan, et al.
Veröffentlicht: (2025)
Penalty shootouts are tough, but the alternating order is fair
von: Vollmer, Silvan, et al.
Veröffentlicht: (2023)
von: Vollmer, Silvan, et al.
Veröffentlicht: (2023)
Darts Analysis
von: Makhamra, Ayham, et al.
Veröffentlicht: (2025)
von: Makhamra, Ayham, et al.
Veröffentlicht: (2025)
There to care; not to kill: medical settings, statistics and wrongful convictions
von: Gill, Richard D.
Veröffentlicht: (2026)
von: Gill, Richard D.
Veröffentlicht: (2026)
Taskmaster Deconstructed: A Quantitative Look at Tension, Volatility, and Viewer Ratings
von: Silver, David H.
Veröffentlicht: (2025)
von: Silver, David H.
Veröffentlicht: (2025)
Modelling handball outcomes using univariate and bivariate approaches
von: Karlis, Dimitris, et al.
Veröffentlicht: (2024)
von: Karlis, Dimitris, et al.
Veröffentlicht: (2024)
Statistical Measures for Explainable Aspect-Based Sentiment Analysis: A Case Study on Environmental Discourse in Reddit
von: Stracqualursi, Luisa, et al.
Veröffentlicht: (2026)
von: Stracqualursi, Luisa, et al.
Veröffentlicht: (2026)
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
von: Ashuach, Tomer, et al.
Veröffentlicht: (2025)
von: Ashuach, Tomer, et al.
Veröffentlicht: (2025)
Towards Greater Leverage: Scaling Laws for Efficient Mixture-of-Experts Language Models
von: Tian, Changxin, et al.
Veröffentlicht: (2025)
von: Tian, Changxin, et al.
Veröffentlicht: (2025)
Mapping the Web of Science, a large-scale graph and text-based dataset with LLM embeddings
von: Kunt, Tim, et al.
Veröffentlicht: (2026)
von: Kunt, Tim, et al.
Veröffentlicht: (2026)
Loss Functions for Detecting Outliers in Panel Data
von: Coleman, Charles D., et al.
Veröffentlicht: (2025)
von: Coleman, Charles D., et al.
Veröffentlicht: (2025)
A model and method for analyzing the precision of binary measurement methods based on beta-binomial distributions, and related statistical tests
von: Takeshita, Jun-ichi, et al.
Veröffentlicht: (2020)
von: Takeshita, Jun-ichi, et al.
Veröffentlicht: (2020)
Artificial intelligence and downscaling global climate model future projections
von: Benestad, Rasmus E.
Veröffentlicht: (2026)
von: Benestad, Rasmus E.
Veröffentlicht: (2026)
Statistical Challenges in Analyzing Migrant Backgrounds Among University Students: a Case Study from Italy
von: Giammei, Lorenzo, et al.
Veröffentlicht: (2025)
von: Giammei, Lorenzo, et al.
Veröffentlicht: (2025)
Towards Explainable Automated Data Quality Enhancement without Domain Knowledge
von: Sarr, Djibril
Veröffentlicht: (2024)
von: Sarr, Djibril
Veröffentlicht: (2024)
GIM: Evaluating models via tasks that integrate multiple cognitive domains
von: Patel, Rohit, et al.
Veröffentlicht: (2026)
von: Patel, Rohit, et al.
Veröffentlicht: (2026)
Directional Asymmetry in Edge BasedSpatial Models via a Skew Normal Prior
von: Cruz-Reyes, Danna L., et al.
Veröffentlicht: (2026)
von: Cruz-Reyes, Danna L., et al.
Veröffentlicht: (2026)
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
von: Peters, Sydney, et al.
Veröffentlicht: (2025)
von: Peters, Sydney, et al.
Veröffentlicht: (2025)
The Likelihood Ratio Wall: Structural Limits on Accurate Risk Assessment for Rare Violence
von: Pollanen, Marco
Veröffentlicht: (2026)
von: Pollanen, Marco
Veröffentlicht: (2026)
Emergent Lexical Semantics in Neural Language Models: Testing Martin's Law on LLM-Generated Text
von: Kugler, Kai
Veröffentlicht: (2025)
von: Kugler, Kai
Veröffentlicht: (2025)
Towards Fundamental Language Models: Does Linguistic Competence Scale with Model Size?
von: Collado-Montañez, Jaime, et al.
Veröffentlicht: (2025)
von: Collado-Montañez, Jaime, et al.
Veröffentlicht: (2025)
Hierarchical Bayesian Knowledge Tracing in Undergraduate Engineering Education
von: Sun, Yiwei
Veröffentlicht: (2025)
von: Sun, Yiwei
Veröffentlicht: (2025)
Masked by Consensus: Disentangling Privileged Knowledge in LLM Correctness
von: Ashuach, Tomer, et al.
Veröffentlicht: (2026)
von: Ashuach, Tomer, et al.
Veröffentlicht: (2026)
AI chatbots versus human healthcare professionals: a systematic review and meta-analysis of empathy in patient care
von: Howcroft, Alastair, et al.
Veröffentlicht: (2026)
von: Howcroft, Alastair, et al.
Veröffentlicht: (2026)
Beyond Human Judgment: A Bayesian Evaluation of LLMs' Moral Values Understanding
von: Skorski, Maciej, et al.
Veröffentlicht: (2025)
von: Skorski, Maciej, et al.
Veröffentlicht: (2025)
Low-Resource Court Judgment Summarization for Common Law Systems
von: Liu, Shuaiqi, et al.
Veröffentlicht: (2024)
von: Liu, Shuaiqi, et al.
Veröffentlicht: (2024)
Beyond Rating: A Comprehensive Evaluation and Benchmark for AI Reviews
von: Li, Bowen, et al.
Veröffentlicht: (2026)
von: Li, Bowen, et al.
Veröffentlicht: (2026)
RomanLens: The Role Of Latent Romanization In Multilinguality In LLMs
von: Saji, Alan, et al.
Veröffentlicht: (2025)
von: Saji, Alan, et al.
Veröffentlicht: (2025)
Penalized Sparse Covariance Regression with High Dimensional Covariates
von: Gao, Yuan, et al.
Veröffentlicht: (2024)
von: Gao, Yuan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Word Embedding Techniques for Classification of Star Ratings
von: Abdelmotaleb, Hesham, et al.
Veröffentlicht: (2025) -
From Noisy News Sentiment Scores to Interpretable Temporal Dynamics: A Bayesian State-Space Model
von: Casals, Ian Carbó
Veröffentlicht: (2026) -
A Structural Text-Based Scaling Model for Analyzing Political Discourse
von: Vávra, Jan, et al.
Veröffentlicht: (2024) -
Claim Automation using Large Language Model
von: Mo, Zhengda, et al.
Veröffentlicht: (2026) -
Pre-trained Models Perform the Best When Token Distributions Follow Zipf's Law
von: He, Yanjin, et al.
Veröffentlicht: (2025)