Tracking linguistic information in transformer-based sentence embeddings through targeted sparsification
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Nastase, Vivi, Merlo, Paola |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Are there identifiable structural parts in the sentence embedding whole?
von: Nastase, Vivi, et al.
Veröffentlicht: (2024)
von: Nastase, Vivi, et al.
Veröffentlicht: (2024)
Testing the assumptions about the geometry of sentence embedding spaces: the cosine measure need not apply
von: Nastase, Vivi, et al.
Veröffentlicht: (2025)
von: Nastase, Vivi, et al.
Veröffentlicht: (2025)
Exploring syntactic information in sentence embeddings through multilingual subject-verb agreement
von: Nastase, Vivi, et al.
Veröffentlicht: (2024)
von: Nastase, Vivi, et al.
Veröffentlicht: (2024)
Exploring Italian sentence embeddings properties through multi-tasking
von: Nastase, Vivi, et al.
Veröffentlicht: (2024)
von: Nastase, Vivi, et al.
Veröffentlicht: (2024)
Exploring State Tracking Capabilities of Large Language Models
von: Rezaee, Kiamehr, et al.
Veröffentlicht: (2025)
von: Rezaee, Kiamehr, et al.
Veröffentlicht: (2025)
The Unlikely Duel: Evaluating Creative Writing in LLMs through a Unique Scenario
von: Gómez-Rodríguez, Carlos, et al.
Veröffentlicht: (2024)
von: Gómez-Rodríguez, Carlos, et al.
Veröffentlicht: (2024)
Evaluating Input Feature Explanations through a Unified Diagnostic Evaluation Framework
von: Sun, Jingyi, et al.
Veröffentlicht: (2024)
von: Sun, Jingyi, et al.
Veröffentlicht: (2024)
GATE: Graph-based Adaptive Tool Evolution Across Diverse Tasks
von: Luo, Jianwen, et al.
Veröffentlicht: (2025)
von: Luo, Jianwen, et al.
Veröffentlicht: (2025)
Accurate Retraining-free Pruning for Pretrained Encoder-based Language Models
von: Park, Seungcheol, et al.
Veröffentlicht: (2023)
von: Park, Seungcheol, et al.
Veröffentlicht: (2023)
Prior-based Noisy Text Data Filtering: Fast and Strong Alternative For Perplexity
von: Seo, Yeongbin, et al.
Veröffentlicht: (2025)
von: Seo, Yeongbin, et al.
Veröffentlicht: (2025)
PaperAudit-Bench: Benchmarking Error Detection in Research Papers for Critical Automated Peer Review
von: Tu, Songjun, et al.
Veröffentlicht: (2026)
von: Tu, Songjun, et al.
Veröffentlicht: (2026)
Co-NAML-LSTUR: A Combined Model with Attentive Multi-View Learning and Long- and Short-term User Representations for News Recommendation
von: Nguyen, Minh Hoang, et al.
Veröffentlicht: (2025)
von: Nguyen, Minh Hoang, et al.
Veröffentlicht: (2025)
The Knesset Corpus: An Annotated Corpus of Hebrew Parliamentary Proceedings
von: Goldin, Gili, et al.
Veröffentlicht: (2024)
von: Goldin, Gili, et al.
Veröffentlicht: (2024)
Towards Effective and Efficient Continual Pre-training of Large Language Models
von: Chen, Jie, et al.
Veröffentlicht: (2024)
von: Chen, Jie, et al.
Veröffentlicht: (2024)
Direct Large Language Model Alignment Through Self-Rewarding Contrastive Prompt Distillation
von: Liu, Aiwei, et al.
Veröffentlicht: (2024)
von: Liu, Aiwei, et al.
Veröffentlicht: (2024)
SentiCSE: A Sentiment-aware Contrastive Sentence Embedding Framework with Sentiment-guided Textual Similarity
von: Kim, Jaemin, et al.
Veröffentlicht: (2024)
von: Kim, Jaemin, et al.
Veröffentlicht: (2024)
The Superalignment of Superhuman Intelligence with Large Language Models
von: Huang, Minlie, et al.
Veröffentlicht: (2024)
von: Huang, Minlie, et al.
Veröffentlicht: (2024)
Train-Attention: Meta-Learning Where to Focus in Continual Knowledge Learning
von: Seo, Yeongbin, et al.
Veröffentlicht: (2024)
von: Seo, Yeongbin, et al.
Veröffentlicht: (2024)
Experimentation in Content Moderation using RWKV
von: Yildirim, Umut, et al.
Veröffentlicht: (2024)
von: Yildirim, Umut, et al.
Veröffentlicht: (2024)
Branching Narratives: Character Decision Points Detection
von: Tikhonov, Alexey
Veröffentlicht: (2024)
von: Tikhonov, Alexey
Veröffentlicht: (2024)
Dancing in the syntax forest: fast, accurate and explainable sentiment analysis with SALSA
von: Gómez-Rodríguez, Carlos, et al.
Veröffentlicht: (2024)
von: Gómez-Rodríguez, Carlos, et al.
Veröffentlicht: (2024)
On the Robustness of Document-Level Relation Extraction Models to Entity Name Variations
von: Meng, Shiao, et al.
Veröffentlicht: (2024)
von: Meng, Shiao, et al.
Veröffentlicht: (2024)
Can LLMs Compute with Reasons?
von: Sandilya, Harshit, et al.
Veröffentlicht: (2024)
von: Sandilya, Harshit, et al.
Veröffentlicht: (2024)
A Survey on Natural Language Counterfactual Generation
von: Wang, Yongjie, et al.
Veröffentlicht: (2024)
von: Wang, Yongjie, et al.
Veröffentlicht: (2024)
Large Language Models(LLMs) on Tabular Data: Prediction, Generation, and Understanding -- A Survey
von: Fang, Xi, et al.
Veröffentlicht: (2024)
von: Fang, Xi, et al.
Veröffentlicht: (2024)
TIS-DPO: Token-level Importance Sampling for Direct Preference Optimization With Estimated Weights
von: Liu, Aiwei, et al.
Veröffentlicht: (2024)
von: Liu, Aiwei, et al.
Veröffentlicht: (2024)
WaterSeeker: Pioneering Efficient Detection of Watermarked Segments in Large Documents
von: Pan, Leyi, et al.
Veröffentlicht: (2024)
von: Pan, Leyi, et al.
Veröffentlicht: (2024)
Distilling Large Language Models for Efficient Clinical Information Extraction
von: Vedula, Karthik S., et al.
Veröffentlicht: (2024)
von: Vedula, Karthik S., et al.
Veröffentlicht: (2024)
MELoRA: Mini-Ensemble Low-Rank Adapters for Parameter-Efficient Fine-Tuning
von: Ren, Pengjie, et al.
Veröffentlicht: (2024)
von: Ren, Pengjie, et al.
Veröffentlicht: (2024)
Evaluating Pixel Language Models on Non-Standardized Languages
von: Muñoz-Ortiz, Alberto, et al.
Veröffentlicht: (2024)
von: Muñoz-Ortiz, Alberto, et al.
Veröffentlicht: (2024)
From Brazilian Portuguese to European Portuguese
von: Sanches, João, et al.
Veröffentlicht: (2024)
von: Sanches, João, et al.
Veröffentlicht: (2024)
Measuring text summarization factuality using atomic facts entailment metrics in the context of retrieval augmented generation
von: Kriman, N. E.
Veröffentlicht: (2024)
von: Kriman, N. E.
Veröffentlicht: (2024)
Setting Standards in Turkish NLP: TR-MMLU for Large Language Model Evaluation
von: Bayram, M. Ali, et al.
Veröffentlicht: (2024)
von: Bayram, M. Ali, et al.
Veröffentlicht: (2024)
Math Natural Language Inference: this should be easy!
von: de Paiva, Valeria, et al.
Veröffentlicht: (2025)
von: de Paiva, Valeria, et al.
Veröffentlicht: (2025)
Fast Quiet-STaR: Thinking Without Thought Tokens
von: Huang, Wei, et al.
Veröffentlicht: (2025)
von: Huang, Wei, et al.
Veröffentlicht: (2025)
New Skills or Sharper Primitives? A Probabilistic Perspective on the Emergence of Reasoning in RLVR
von: Wang, Zhilin, et al.
Veröffentlicht: (2026)
von: Wang, Zhilin, et al.
Veröffentlicht: (2026)
d-TreeRPO: Towards More Reliable Policy Optimization for Diffusion Language Models
von: Pan, Leyi, et al.
Veröffentlicht: (2025)
von: Pan, Leyi, et al.
Veröffentlicht: (2025)
An Unforgeable Publicly Verifiable Watermark for Large Language Models
von: Liu, Aiwei, et al.
Veröffentlicht: (2023)
von: Liu, Aiwei, et al.
Veröffentlicht: (2023)
Exploiting Pre-trained Encoder-Decoder Transformers for Sequence-to-Sequence Constituent Parsing
von: Fernández-González, Daniel, et al.
Veröffentlicht: (2026)
von: Fernández-González, Daniel, et al.
Veröffentlicht: (2026)
Parametric Social Identity Injection and Diversification in Public Opinion Simulation
von: Wang, Hexi, et al.
Veröffentlicht: (2026)
von: Wang, Hexi, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Are there identifiable structural parts in the sentence embedding whole?
von: Nastase, Vivi, et al.
Veröffentlicht: (2024) -
Testing the assumptions about the geometry of sentence embedding spaces: the cosine measure need not apply
von: Nastase, Vivi, et al.
Veröffentlicht: (2025) -
Exploring syntactic information in sentence embeddings through multilingual subject-verb agreement
von: Nastase, Vivi, et al.
Veröffentlicht: (2024) -
Exploring Italian sentence embeddings properties through multi-tasking
von: Nastase, Vivi, et al.
Veröffentlicht: (2024) -
Exploring State Tracking Capabilities of Large Language Models
von: Rezaee, Kiamehr, et al.
Veröffentlicht: (2025)