Thinking like a CHEMIST: Combined Heterogeneous Embedding Model Integrating Structure and Tokens
Fuente:
arXiv
Saved in:
| Main Authors: | Rekut, Nikolai, Orlov, Alexey, Ziu, Klea, Starykh, Elizaveta, Takac, Martin, Beznosikov, Aleksandr |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MirrorCheck: Efficient Adversarial Defense for Vision-Language Models
by: Fares, Samar, et al.
Published: (2024)
by: Fares, Samar, et al.
Published: (2024)
$ψ$DAG: Projected Stochastic Approximation Iteration for DAG Structure Learning
by: Ziu, Klea, et al.
Published: (2024)
by: Ziu, Klea, et al.
Published: (2024)
Random-reshuffled SARAH does not need a full gradient computations
by: Beznosikov, Aleksandr, et al.
Published: (2021)
by: Beznosikov, Aleksandr, et al.
Published: (2021)
WaveSSM: Multiscale State-Space Models for Non-stationary Signal Attention
by: Solozabal, Ruben, et al.
Published: (2026)
by: Solozabal, Ruben, et al.
Published: (2026)
HARP: Hadamard-Preconditioned Adaptive Rotation Processor for Extreme LLM Quantization
by: Zagitov, Artur, et al.
Published: (2026)
by: Zagitov, Artur, et al.
Published: (2026)
Hierarchical Mixture-of-Experts with Two-Stage Optimization
by: Molodtsov, Gleb, et al.
Published: (2026)
by: Molodtsov, Gleb, et al.
Published: (2026)
Similarity, Compression and Local Steps: Three Pillars of Efficient Communications for Distributed Variational Inequalities
by: Beznosikov, Aleksandr, et al.
Published: (2023)
by: Beznosikov, Aleksandr, et al.
Published: (2023)
Extreme Low-Bit Inference in Reasoning Models: Failure Modes and Targeted Recovery
by: Alimaskina, Ekaterina, et al.
Published: (2026)
by: Alimaskina, Ekaterina, et al.
Published: (2026)
FRUGAL: Memory-Efficient Optimization by Reducing State Overhead for Scalable Training
by: Zmushko, Philip, et al.
Published: (2024)
by: Zmushko, Philip, et al.
Published: (2024)
A theoretical framework for self-supervised contrastive learning for continuous dependent data
by: Marusov, Alexander, et al.
Published: (2025)
by: Marusov, Alexander, et al.
Published: (2025)
Think Then Embed: Generative Context Improves Multimodal Embedding
by: Cui, Xuanming, et al.
Published: (2025)
by: Cui, Xuanming, et al.
Published: (2025)
Rethinking Thinking Tokens: LLMs as Improvement Operators
by: Madaan, Lovish, et al.
Published: (2025)
by: Madaan, Lovish, et al.
Published: (2025)
Recurrent Action Transformer with Memory
by: Cherepanov, Egor, et al.
Published: (2023)
by: Cherepanov, Egor, et al.
Published: (2023)
Uncovering the Spectral Bias in Diagonal State Space Models
by: Solozabal, Ruben, et al.
Published: (2025)
by: Solozabal, Ruben, et al.
Published: (2025)
UPath: Universal Planner Across Topological Heterogeneity For Grid-Based Pathfinding
by: Ananikian, Aleksandr, et al.
Published: (2026)
by: Ananikian, Aleksandr, et al.
Published: (2026)
The Ky Fan Norms and Beyond: Dual Norms and Combinations for Matrix Optimization
by: Kravatskiy, Alexey, et al.
Published: (2025)
by: Kravatskiy, Alexey, et al.
Published: (2025)
Embedding-Aware Feature Discovery: Bridging Latent Representations and Interpretable Features in Event Sequences
by: Sakhno, Artem, et al.
Published: (2026)
by: Sakhno, Artem, et al.
Published: (2026)
Revisiting Tree Search for LLMs: Gumbel and Sequential Halving for Budget-Scalable Reasoning
by: Ugadiarov, Leonid, et al.
Published: (2026)
by: Ugadiarov, Leonid, et al.
Published: (2026)
CAMAR: Continuous Actions Multi-Agent Routing
by: Pshenitsyn, Artem, et al.
Published: (2025)
by: Pshenitsyn, Artem, et al.
Published: (2025)
Where Does Warm-Up Come From? Adaptive Scheduling for Norm-Constrained Optimizers
by: Riabinin, Artem, et al.
Published: (2026)
by: Riabinin, Artem, et al.
Published: (2026)
Dynamic Thinking-Token Selection for Efficient Reasoning in Large Reasoning Models
by: Guo, Zhenyuan, et al.
Published: (2026)
by: Guo, Zhenyuan, et al.
Published: (2026)
Think before you speak: Training Language Models With Pause Tokens
by: Goyal, Sachin, et al.
Published: (2023)
by: Goyal, Sachin, et al.
Published: (2023)
Dynamics of Moral Behavior in Heterogeneous Populations of Learning Agents
by: Tennant, Elizaveta, et al.
Published: (2024)
by: Tennant, Elizaveta, et al.
Published: (2024)
From Muscle Bursts to Motor Intent: Self-Supervised Token Modeling for Heterogeneous EMG
by: Huang, Zhenghao, et al.
Published: (2026)
by: Huang, Zhenghao, et al.
Published: (2026)
On the Bias of Next-Token Predictors Toward Systematically Inefficient Reasoning: A Shortest-Path Case Study
by: Alberghi, Riccardo, et al.
Published: (2025)
by: Alberghi, Riccardo, et al.
Published: (2025)
JTok: On Token Embedding as another Axis of Scaling Law via Joint Token Self-modulation
by: Yang, Yebin, et al.
Published: (2026)
by: Yang, Yebin, et al.
Published: (2026)
Latent Reasoning in TRMs is Secretly a Policy Improvement Operator
by: Asadulaev, Arip, et al.
Published: (2025)
by: Asadulaev, Arip, et al.
Published: (2025)
ELMUR: External Layer Memory with Update/Rewrite for Long-Horizon RL Problems
by: Cherepanov, Egor, et al.
Published: (2025)
by: Cherepanov, Egor, et al.
Published: (2025)
Probing the Embedding Space of Transformers via Minimal Token Perturbations
by: Conti, Eddie, et al.
Published: (2025)
by: Conti, Eddie, et al.
Published: (2025)
Memory Retention Is Not Enough to Master Memory Tasks in Reinforcement Learning
by: Shchendrigin, Oleg, et al.
Published: (2026)
by: Shchendrigin, Oleg, et al.
Published: (2026)
Re:Frame -- Retrieving Experience From Associative Memory
by: Zelezetsky, Daniil, et al.
Published: (2025)
by: Zelezetsky, Daniil, et al.
Published: (2025)
Think like a Scientist: Physics-guided LLM Agent for Equation Discovery
by: Yang, Jianke, et al.
Published: (2026)
by: Yang, Jianke, et al.
Published: (2026)
Dywave: Event-Aligned Dynamic Tokenization for Heterogeneous IoT Sensing Signals
by: Kimura, Tomoyoshi, et al.
Published: (2026)
by: Kimura, Tomoyoshi, et al.
Published: (2026)
The AI Data Scientist
by: Akimov, Farkhad, et al.
Published: (2025)
by: Akimov, Farkhad, et al.
Published: (2025)
Zero-Shot Off-Policy Learning
by: Asadulaev, Arip, et al.
Published: (2026)
by: Asadulaev, Arip, et al.
Published: (2026)
Y-Shaped Generative Flows
by: Asadulaev, Arip, et al.
Published: (2025)
by: Asadulaev, Arip, et al.
Published: (2025)
LookPlanGraph: Embodied Instruction Following Method with VLM Graph Augmentation
by: Onishchenko, Anatoly O., et al.
Published: (2025)
by: Onishchenko, Anatoly O., et al.
Published: (2025)
One Token Embedding Is Enough to Deadlock Your Large Reasoning Model
by: Zhang, Mohan, et al.
Published: (2025)
by: Zhang, Mohan, et al.
Published: (2025)
Think Clearly: Improving Reasoning via Redundant Token Pruning
by: Choi, Daewon, et al.
Published: (2025)
by: Choi, Daewon, et al.
Published: (2025)
Multiplex Thinking: Reasoning via Token-wise Branch-and-Merge
by: Tang, Yao, et al.
Published: (2026)
by: Tang, Yao, et al.
Published: (2026)
Similar Items
-
MirrorCheck: Efficient Adversarial Defense for Vision-Language Models
by: Fares, Samar, et al.
Published: (2024) -
$ψ$DAG: Projected Stochastic Approximation Iteration for DAG Structure Learning
by: Ziu, Klea, et al.
Published: (2024) -
Random-reshuffled SARAH does not need a full gradient computations
by: Beznosikov, Aleksandr, et al.
Published: (2021) -
WaveSSM: Multiscale State-Space Models for Non-stationary Signal Attention
by: Solozabal, Ruben, et al.
Published: (2026) -
HARP: Hadamard-Preconditioned Adaptive Rotation Processor for Extreme LLM Quantization
by: Zagitov, Artur, et al.
Published: (2026)