Patent Representation Learning via Self-supervision
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zuo, You, Gerdes, Kim, de La Clergerie, Eric Villemonte, Sagot, Benoît |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PatentEval: Understanding Errors in Patent Generation
von: Zuo, You, et al.
Veröffentlicht: (2024)
von: Zuo, You, et al.
Veröffentlicht: (2024)
Anisotropy Is Inherent to Self-Attention in Transformers
von: Godey, Nathan, et al.
Veröffentlicht: (2024)
von: Godey, Nathan, et al.
Veröffentlicht: (2024)
On the Scaling Laws of Geographical Representation in Language Models
von: Godey, Nathan, et al.
Veröffentlicht: (2024)
von: Godey, Nathan, et al.
Veröffentlicht: (2024)
Why do small language models underperform? Studying Language Model Saturation via the Softmax Bottleneck
von: Godey, Nathan, et al.
Veröffentlicht: (2024)
von: Godey, Nathan, et al.
Veröffentlicht: (2024)
Biomed-Enriched: A Biomedical Dataset Enriched with LLMs for Pretraining and Extracting Rare and Hidden Content
von: Touchent, Rian, et al.
Veröffentlicht: (2025)
von: Touchent, Rian, et al.
Veröffentlicht: (2025)
Can Character-based Language Models Improve Downstream Task Performance in Low-Resource and Noisy Language Scenarios?
von: Riabi, Arij, et al.
Veröffentlicht: (2021)
von: Riabi, Arij, et al.
Veröffentlicht: (2021)
CamemBERT 2.0: A Smarter French Language Model Aged to Perfection
von: Antoun, Wissam, et al.
Veröffentlicht: (2024)
von: Antoun, Wissam, et al.
Veröffentlicht: (2024)
PaECTER: Patent-level Representation Learning using Citation-informed Transformers
von: Ghosh, Mainak, et al.
Veröffentlicht: (2024)
von: Ghosh, Mainak, et al.
Veröffentlicht: (2024)
Training-Free Cross-Lingual Dysarthria Severity Assessment via Phonological Subspace Analysis in Self-Supervised Speech Representations
von: Muller, Bernard, et al.
Veröffentlicht: (2026)
von: Muller, Bernard, et al.
Veröffentlicht: (2026)
A Confidence-based Acquisition Model for Self-supervised Active Learning and Label Correction
von: van Niekerk, Carel, et al.
Veröffentlicht: (2023)
von: van Niekerk, Carel, et al.
Veröffentlicht: (2023)
Knowledge Graph Reasoning with Self-supervised Reinforcement Learning
von: Ma, Ying, et al.
Veröffentlicht: (2024)
von: Ma, Ying, et al.
Veröffentlicht: (2024)
Polynomial Mixing for Efficient Self-supervised Speech Encoders
von: Feillet, Eva, et al.
Veröffentlicht: (2026)
von: Feillet, Eva, et al.
Veröffentlicht: (2026)
Gaperon: A Peppered English-French Generative Language Model Suite
von: Godey, Nathan, et al.
Veröffentlicht: (2025)
von: Godey, Nathan, et al.
Veröffentlicht: (2025)
Q-Filters: Leveraging QK Geometry for Efficient KV Cache Compression
von: Godey, Nathan, et al.
Veröffentlicht: (2025)
von: Godey, Nathan, et al.
Veröffentlicht: (2025)
On the Relationship Between the Choice of Representation and In-Context Learning
von: Marinescu, Ioana, et al.
Veröffentlicht: (2025)
von: Marinescu, Ioana, et al.
Veröffentlicht: (2025)
$\texttt{PatentAgent}$: Intelligent Agent for Automated Pharmaceutical Patent Analysis
von: Wang, Xin, et al.
Veröffentlicht: (2024)
von: Wang, Xin, et al.
Veröffentlicht: (2024)
Table-r1: Self-supervised and Reinforcement Learning for Program-based Table Reasoning in Small Language Models
von: Jin, Rihui, et al.
Veröffentlicht: (2025)
von: Jin, Rihui, et al.
Veröffentlicht: (2025)
PATENTWRITER: A Benchmarking Study for Patent Drafting with LLMs
von: Shomee, Homaira Huda, et al.
Veröffentlicht: (2025)
von: Shomee, Homaira Huda, et al.
Veröffentlicht: (2025)
ReVISE: Learning to Refine at Test-Time via Intrinsic Self-Verification
von: Lee, Hyunseok, et al.
Veröffentlicht: (2025)
von: Lee, Hyunseok, et al.
Veröffentlicht: (2025)
Graph Contrastive Learning via Cluster-refined Negative Sampling for Semi-supervised Text Classification
von: Ai, Wei, et al.
Veröffentlicht: (2024)
von: Ai, Wei, et al.
Veröffentlicht: (2024)
SNAP-UQ: Self-supervised Next-Activation Prediction for Single-Pass Uncertainty in TinyML
von: Lamaakal, Ismail, et al.
Veröffentlicht: (2025)
von: Lamaakal, Ismail, et al.
Veröffentlicht: (2025)
ReflexiCoder: Teaching Large Language Models to Self-Reflect on Generated Code and Self-Correct It via Reinforcement Learning
von: Jiang, Juyong, et al.
Veröffentlicht: (2026)
von: Jiang, Juyong, et al.
Veröffentlicht: (2026)
Self-Refining Language Model Anonymizers via Adversarial Distillation
von: Kim, Kyuyoung, et al.
Veröffentlicht: (2025)
von: Kim, Kyuyoung, et al.
Veröffentlicht: (2025)
Towards Efficient Active Learning in NLP via Pretrained Representations
von: Vysogorets, Artem, et al.
Veröffentlicht: (2024)
von: Vysogorets, Artem, et al.
Veröffentlicht: (2024)
STARLING: Self-supervised Training of Text-based Reinforcement Learning Agent with Large Language Models
von: Basavatia, Shreyas, et al.
Veröffentlicht: (2024)
von: Basavatia, Shreyas, et al.
Veröffentlicht: (2024)
Enhancing Sindhi Word Segmentation using Subword Representation Learning and Position-aware Self-attention
von: Ali, Wazir, et al.
Veröffentlicht: (2020)
von: Ali, Wazir, et al.
Veröffentlicht: (2020)
S$^2$R: Teaching LLMs to Self-verify and Self-correct via Reinforcement Learning
von: Ma, Ruotian, et al.
Veröffentlicht: (2025)
von: Ma, Ruotian, et al.
Veröffentlicht: (2025)
Patent Language Model Pretraining with ModernBERT
von: Yousefiramandi, Amirhossein, et al.
Veröffentlicht: (2025)
von: Yousefiramandi, Amirhossein, et al.
Veröffentlicht: (2025)
LittleBit: Ultra Low-Bit Quantization via Latent Factorization
von: Lee, Banseok, et al.
Veröffentlicht: (2025)
von: Lee, Banseok, et al.
Veröffentlicht: (2025)
Position Information Emerges in Causal Transformers Without Positional Encodings via Similarity of Nearby Embeddings
von: Zuo, Chunsheng, et al.
Veröffentlicht: (2024)
von: Zuo, Chunsheng, et al.
Veröffentlicht: (2024)
Self-supervised Transformation Learning for Equivariant Representations
von: Yu, Jaemyung, et al.
Veröffentlicht: (2025)
von: Yu, Jaemyung, et al.
Veröffentlicht: (2025)
Learning Task Representations from In-Context Learning
von: Saglam, Baturay, et al.
Veröffentlicht: (2025)
von: Saglam, Baturay, et al.
Veröffentlicht: (2025)
Explicit Learning and the LLM in Machine Translation
von: Marmonier, Malik, et al.
Veröffentlicht: (2025)
von: Marmonier, Malik, et al.
Veröffentlicht: (2025)
DETree: DEtecting Human-AI Collaborative Texts via Tree-Structured Hierarchical Representation Learning
von: He, Yongxin, et al.
Veröffentlicht: (2025)
von: He, Yongxin, et al.
Veröffentlicht: (2025)
Post-Trained MoE Can Skip Half Experts via Self-Distillation
von: Lv, Xingtai, et al.
Veröffentlicht: (2026)
von: Lv, Xingtai, et al.
Veröffentlicht: (2026)
Silence the Judge: Reinforcement Learning with Self-Verifier via Latent Geometric Clustering
von: Zhang, Nonghai, et al.
Veröffentlicht: (2026)
von: Zhang, Nonghai, et al.
Veröffentlicht: (2026)
Learning Harmonized Representations for Speculative Sampling
von: Zhang, Lefan, et al.
Veröffentlicht: (2024)
von: Zhang, Lefan, et al.
Veröffentlicht: (2024)
In-Context Example Selection via Similarity Search Improves Low-Resource Machine Translation
von: Zebaze, Armel, et al.
Veröffentlicht: (2024)
von: Zebaze, Armel, et al.
Veröffentlicht: (2024)
Fair Text Classification via Transferable Representations
von: Leteno, Thibaud, et al.
Veröffentlicht: (2025)
von: Leteno, Thibaud, et al.
Veröffentlicht: (2025)
Learning without training: The implicit dynamics of in-context learning
von: Dherin, Benoit, et al.
Veröffentlicht: (2025)
von: Dherin, Benoit, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
PatentEval: Understanding Errors in Patent Generation
von: Zuo, You, et al.
Veröffentlicht: (2024) -
Anisotropy Is Inherent to Self-Attention in Transformers
von: Godey, Nathan, et al.
Veröffentlicht: (2024) -
On the Scaling Laws of Geographical Representation in Language Models
von: Godey, Nathan, et al.
Veröffentlicht: (2024) -
Why do small language models underperform? Studying Language Model Saturation via the Softmax Bottleneck
von: Godey, Nathan, et al.
Veröffentlicht: (2024) -
Biomed-Enriched: A Biomedical Dataset Enriched with LLMs for Pretraining and Extracting Rare and Hidden Content
von: Touchent, Rian, et al.
Veröffentlicht: (2025)