ArXiv-to-Model: A Practical Study of Scientific LM Training
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Gupta, Anuj |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Multimodal ArXiv: A Dataset for Improving Scientific Comprehension of Large Vision-Language Models
von: Li, Lei, et al.
Veröffentlicht: (2024)
von: Li, Lei, et al.
Veröffentlicht: (2024)
LLM-Generated or Human-Written? Comparing Review and Non-Review Papers on ArXiv
von: Elazar, Yanai, et al.
Veröffentlicht: (2026)
von: Elazar, Yanai, et al.
Veröffentlicht: (2026)
To ArXiv or not to ArXiv: A Study Quantifying Pros and Cons of Posting Preprints Online
von: Rastogi, Charvi, et al.
Veröffentlicht: (2022)
von: Rastogi, Charvi, et al.
Veröffentlicht: (2022)
Scientific Statement Classification over arXiv.org
von: Ginev, Deyan, et al.
Veröffentlicht: (2019)
von: Ginev, Deyan, et al.
Veröffentlicht: (2019)
aiXiv: A Next-Generation Open Access Ecosystem for Scientific Discovery Generated by AI Scientists
von: Zhang, Pengsong, et al.
Veröffentlicht: (2025)
von: Zhang, Pengsong, et al.
Veröffentlicht: (2025)
DeepXiv-SDK: An Agentic Data Interface for Scientific Literature
von: Qian, Hongjin, et al.
Veröffentlicht: (2026)
von: Qian, Hongjin, et al.
Veröffentlicht: (2026)
ArXivBench: When You Should Avoid Using ChatGPT for Academic Writing
von: Li, Ning, et al.
Veröffentlicht: (2025)
von: Li, Ning, et al.
Veröffentlicht: (2025)
BioMedLM: A 2.7B Parameter Language Model Trained On Biomedical Text
von: Bolton, Elliot, et al.
Veröffentlicht: (2024)
von: Bolton, Elliot, et al.
Veröffentlicht: (2024)
Iterative Auto-Annotation for Scientific Named Entity Recognition Using BERT-Based Models
von: Gupta, Kartik
Veröffentlicht: (2025)
von: Gupta, Kartik
Veröffentlicht: (2025)
Effects of Research Paper Promotion via ArXiv and X
von: Bagchi, Chhandak, et al.
Veröffentlicht: (2024)
von: Bagchi, Chhandak, et al.
Veröffentlicht: (2024)
EvoLM: In Search of Lost Language Model Training Dynamics
von: Qi, Zhenting, et al.
Veröffentlicht: (2025)
von: Qi, Zhenting, et al.
Veröffentlicht: (2025)
LM2: Large Memory Models
von: Kang, Jikun, et al.
Veröffentlicht: (2025)
von: Kang, Jikun, et al.
Veröffentlicht: (2025)
LOGIC-LM++: Multi-Step Refinement for Symbolic Formulations
von: Kirtania, Shashank, et al.
Veröffentlicht: (2024)
von: Kirtania, Shashank, et al.
Veröffentlicht: (2024)
TeachLM: Post-Training LLMs for Education Using Authentic Learning Data
von: Perczel, Janos, et al.
Veröffentlicht: (2025)
von: Perczel, Janos, et al.
Veröffentlicht: (2025)
UrduLM: A Resource-Efficient Monolingual Urdu Language Model
von: Ali, Syed Muhammad, et al.
Veröffentlicht: (2026)
von: Ali, Syed Muhammad, et al.
Veröffentlicht: (2026)
Locale-Conditioned Few-Shot Prompting Mitigates Demonstration Regurgitation in On-Device PII Substitution with Small Language Models
von: Sadani, Anuj, et al.
Veröffentlicht: (2026)
von: Sadani, Anuj, et al.
Veröffentlicht: (2026)
ArCHer: Training Language Model Agents via Hierarchical Multi-Turn RL
von: Zhou, Yifei, et al.
Veröffentlicht: (2024)
von: Zhou, Yifei, et al.
Veröffentlicht: (2024)
Xmodel-LM Technical Report
von: Wang, Yichuan, et al.
Veröffentlicht: (2024)
von: Wang, Yichuan, et al.
Veröffentlicht: (2024)
CogLM: Tracking Cognitive Development of Large Language Models
von: Wang, Xinglin, et al.
Veröffentlicht: (2024)
von: Wang, Xinglin, et al.
Veröffentlicht: (2024)
PonderLM: Pretraining Language Models to Ponder in Continuous Space
von: Zeng, Boyi, et al.
Veröffentlicht: (2025)
von: Zeng, Boyi, et al.
Veröffentlicht: (2025)
Foundation CAN LM: A Pretrained Language Model For Automotive CAN Data
von: Esashi, Akiharu, et al.
Veröffentlicht: (2026)
von: Esashi, Akiharu, et al.
Veröffentlicht: (2026)
JudgeLM: Fine-tuned Large Language Models are Scalable Judges
von: Zhu, Lianghui, et al.
Veröffentlicht: (2023)
von: Zhu, Lianghui, et al.
Veröffentlicht: (2023)
REFINE-LM: Mitigating Language Model Stereotypes via Reinforcement Learning
von: Qureshi, Rameez, et al.
Veröffentlicht: (2024)
von: Qureshi, Rameez, et al.
Veröffentlicht: (2024)
InternLM2 Technical Report
von: Cai, Zheng, et al.
Veröffentlicht: (2024)
von: Cai, Zheng, et al.
Veröffentlicht: (2024)
Ar-Spider: Text-to-SQL in Arabic
von: Almohaimeed, Saleh, et al.
Veröffentlicht: (2024)
von: Almohaimeed, Saleh, et al.
Veröffentlicht: (2024)
KALE-LM-Chem: Vision and Practice Toward an AI Brain for Chemistry
von: Dai, Weichen, et al.
Veröffentlicht: (2024)
von: Dai, Weichen, et al.
Veröffentlicht: (2024)
Identifying the Risks of LM Agents with an LM-Emulated Sandbox
von: Ruan, Yangjun, et al.
Veröffentlicht: (2023)
von: Ruan, Yangjun, et al.
Veröffentlicht: (2023)
Wayfinding through the AI wilderness: Mapping rhetorics of ChatGPT prompt writing on X (formerly Twitter) to promote critical AI literacies
von: Gupta, Anuj, et al.
Veröffentlicht: (2025)
von: Gupta, Anuj, et al.
Veröffentlicht: (2025)
FlowLM: Few-Step Language Modeling via Diffusion-to-Flow Adaptation
von: Zhang, Runzhe, et al.
Veröffentlicht: (2026)
von: Zhang, Runzhe, et al.
Veröffentlicht: (2026)
KG-BiLM: Knowledge Graph Embedding via Bidirectional Language Models
von: Chen, Zirui, et al.
Veröffentlicht: (2025)
von: Chen, Zirui, et al.
Veröffentlicht: (2025)
LM-Infinite: Zero-Shot Extreme Length Generalization for Large Language Models
von: Han, Chi, et al.
Veröffentlicht: (2023)
von: Han, Chi, et al.
Veröffentlicht: (2023)
ShishuLM : Achieving Optimal and Efficient Parameterization with Low Attention Transformer Models
von: Kumar, Shivanshu, et al.
Veröffentlicht: (2025)
von: Kumar, Shivanshu, et al.
Veröffentlicht: (2025)
$\texttt{LM}^\texttt{2}$: A Simple Society of Language Models Solves Complex Reasoning
von: Juneja, Gurusha, et al.
Veröffentlicht: (2024)
von: Juneja, Gurusha, et al.
Veröffentlicht: (2024)
From Euler to Today: Universal Mathematical Fallibility A Large-Scale Computational Analysis of Errors in ArXiv Papers
von: Rivin, Igor
Veröffentlicht: (2025)
von: Rivin, Igor
Veröffentlicht: (2025)
RexBERT: Context Specialized Bidirectional Encoders for E-commerce
von: Bajaj, Rahul, et al.
Veröffentlicht: (2026)
von: Bajaj, Rahul, et al.
Veröffentlicht: (2026)
ReaLM: Residual Quantization Bridging Knowledge Graph Embeddings and Large Language Models
von: Guo, Wenbin, et al.
Veröffentlicht: (2025)
von: Guo, Wenbin, et al.
Veröffentlicht: (2025)
B-cos LM: Efficiently Transforming Pre-trained Language Models for Improved Explainability
von: Wang, Yifan, et al.
Veröffentlicht: (2025)
von: Wang, Yifan, et al.
Veröffentlicht: (2025)
TeacherLM: Teaching to Fish Rather Than Giving the Fish, Language Modeling Likewise
von: He, Nan, et al.
Veröffentlicht: (2023)
von: He, Nan, et al.
Veröffentlicht: (2023)
Fineweb-Edu-Ar: Machine-translated Corpus to Support Arabic Small Language Models
von: Alrashed, Sultan, et al.
Veröffentlicht: (2024)
von: Alrashed, Sultan, et al.
Veröffentlicht: (2024)
PulseLM: A Foundation Dataset and Benchmark for PPG-Text Learning
von: Pham, Hung Manh, et al.
Veröffentlicht: (2026)
von: Pham, Hung Manh, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Multimodal ArXiv: A Dataset for Improving Scientific Comprehension of Large Vision-Language Models
von: Li, Lei, et al.
Veröffentlicht: (2024) -
LLM-Generated or Human-Written? Comparing Review and Non-Review Papers on ArXiv
von: Elazar, Yanai, et al.
Veröffentlicht: (2026) -
To ArXiv or not to ArXiv: A Study Quantifying Pros and Cons of Posting Preprints Online
von: Rastogi, Charvi, et al.
Veröffentlicht: (2022) -
Scientific Statement Classification over arXiv.org
von: Ginev, Deyan, et al.
Veröffentlicht: (2019) -
aiXiv: A Next-Generation Open Access Ecosystem for Scientific Discovery Generated by AI Scientists
von: Zhang, Pengsong, et al.
Veröffentlicht: (2025)