Counting Cycles with Deepseek
Fuente:
arXiv
Saved in:
| Main Authors: | Jin, Jiashun, Ke, Tracy, Sui, Bingcheng, Wang, Zhenggang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Comparison of DeepSeek and Other LLMs
by: Gao, Tianchen, et al.
Published: (2025)
by: Gao, Tianchen, et al.
Published: (2025)
Knowledge Graph-Driven Retrieval-Augmented Generation: Integrating Deepseek-R1 with Weaviate for Advanced Chatbot Applications
by: Lecu, Alexandru, et al.
Published: (2025)
by: Lecu, Alexandru, et al.
Published: (2025)
A Status Quo Investigation of Large Language Models towards Cost-Effective CFD Automation with OpenFOAMGPT: ChatGPT vs. Qwen vs. Deepseek
by: Wang, Wenkang, et al.
Published: (2025)
by: Wang, Wenkang, et al.
Published: (2025)
Can generative AI figure out figurative language? The influence of idioms on essay scoring by ChatGPT, Gemini, and Deepseek
by: Oğuz, Enis
Published: (2025)
by: Oğuz, Enis
Published: (2025)
A comprehensive study of LLM-based argument classification: from LLAMA through GPT-4o to Deepseek-R1
by: Pietroń, Marcin, et al.
Published: (2025)
by: Pietroń, Marcin, et al.
Published: (2025)
A Method for the Architecture of a Medical Vertical Large Language Model Based on Deepseek R1
by: Zhang, Mingda, et al.
Published: (2025)
by: Zhang, Mingda, et al.
Published: (2025)
Stochastic Attention: Connectome-Inspired Randomized Routing for Expressive Linear-Time Attention
by: Jin, Zehao, et al.
Published: (2026)
by: Jin, Zehao, et al.
Published: (2026)
CtrlRAG: Black-box Document Poisoning Attacks for Retrieval-Augmented Generation of Large Language Models
by: Sui, Runqi
Published: (2025)
by: Sui, Runqi
Published: (2025)
LLMs Exhibit Significantly Lower Uncertainty in Creative Writing Than Professional Writers
by: Sui, Peiqi
Published: (2026)
by: Sui, Peiqi
Published: (2026)
Machine Translation in the Wild: User Reaction to Xiaohongshu's Built-In Translation Feature
by: He, Sui
Published: (2026)
by: He, Sui
Published: (2026)
DependencyAI: Detecting AI Generated Text through Dependency Parsing
by: Ahmed, Sara, et al.
Published: (2026)
by: Ahmed, Sara, et al.
Published: (2026)
Prompting ChatGPT for Translation: A Comparative Analysis of Translation Brief and Persona Prompts
by: He, Sui
Published: (2024)
by: He, Sui
Published: (2024)
Reducing Hallucinations in Entity Abstract Summarization with Facts-Template Decomposition
by: Zhu, Fangwei, et al.
Published: (2024)
by: Zhu, Fangwei, et al.
Published: (2024)
CoLT: Reasoning with Chain of Latent Tool Calls
by: Zhu, Fangwei, et al.
Published: (2026)
by: Zhu, Fangwei, et al.
Published: (2026)
Counting Like Transformers: Compiling Temporal Counting Logic Into Softmax Transformers
by: Yang, Andy, et al.
Published: (2024)
by: Yang, Andy, et al.
Published: (2024)
CE-RM: A Pointwise Generative Reward Model Optimized via Two-Stage Rollout and Unified Criteria
by: Hu, Xinyu, et al.
Published: (2026)
by: Hu, Xinyu, et al.
Published: (2026)
Choose Your Own Adventure: Interactive E-Books to Improve Word Knowledge and Comprehension Skills
by: Day, Stephanie, et al.
Published: (2024)
by: Day, Stephanie, et al.
Published: (2024)
The Counting Power of Transformers
by: Sälzer, Marco, et al.
Published: (2025)
by: Sälzer, Marco, et al.
Published: (2025)
Not All Demonstration Examples are Equally Beneficial: Reweighting Demonstration Examples for In-Context Learning
by: Yang, Zhe, et al.
Published: (2023)
by: Yang, Zhe, et al.
Published: (2023)
Towards Better RL Training Data Utilization via Second-Order Rollout
by: Yang, Zhe, et al.
Published: (2026)
by: Yang, Zhe, et al.
Published: (2026)
rStar-Coder: Scaling Competitive Code Reasoning with a Large-Scale Verified Dataset
by: Liu, Yifei, et al.
Published: (2025)
by: Liu, Yifei, et al.
Published: (2025)
Predicting Punctuation in Ancient Chinese Texts: A Multi-Layered LSTM and Attention-Based Approach
by: Cai, Tracy, et al.
Published: (2024)
by: Cai, Tracy, et al.
Published: (2024)
Chain-of-Thought Tokens are Computer Program Variables
by: Zhu, Fangwei, et al.
Published: (2025)
by: Zhu, Fangwei, et al.
Published: (2025)
Exploring Activation Patterns of Parameters in Language Models
by: Wang, Yudong, et al.
Published: (2024)
by: Wang, Yudong, et al.
Published: (2024)
Every Step Counts: Step-Level Credit Assignment for Tool-Integrated Text-to-SQL
by: Dai, Yaxun, et al.
Published: (2026)
by: Dai, Yaxun, et al.
Published: (2026)
LLMSR@XLLM25: An Empirical Study of LLM for Structural Reasoning
by: Li, Xinye, et al.
Published: (2025)
by: Li, Xinye, et al.
Published: (2025)
Think Less, Know More: State-Aware Reasoning Compression with Knowledge Guidance for Efficient Reasoning
by: Sui, Yi, et al.
Published: (2026)
by: Sui, Yi, et al.
Published: (2026)
Language Models Encode the Value of Numbers Linearly
by: Zhu, Fangwei, et al.
Published: (2024)
by: Zhu, Fangwei, et al.
Published: (2024)
Counting and Sampling Traces in Regular Languages
by: de Colnet, Alexis, et al.
Published: (2025)
by: de Colnet, Alexis, et al.
Published: (2025)
Network Goodness-of-Fit for the block-model family
by: Jin, Jiashun, et al.
Published: (2025)
by: Jin, Jiashun, et al.
Published: (2025)
Answer Set Counting and its Applications
by: Kabir, Mohimenul
Published: (2025)
by: Kabir, Mohimenul
Published: (2025)
Parameter-Efficient Fine-Tuning via Circular Convolution
by: Chen, Aochuan, et al.
Published: (2024)
by: Chen, Aochuan, et al.
Published: (2024)
Anim-Director: A Large Multimodal Model Powered Agent for Controllable Animation Video Generation
by: Li, Yunxin, et al.
Published: (2024)
by: Li, Yunxin, et al.
Published: (2024)
Making Every Verified Token Count: Adaptive Verification for MoE Speculative Decoding
by: Pan, Lehan, et al.
Published: (2026)
by: Pan, Lehan, et al.
Published: (2026)
Contextual Position Encoding: Learning to Count What's Important
by: Golovneva, Olga, et al.
Published: (2024)
by: Golovneva, Olga, et al.
Published: (2024)
Conversation for Non-verifiable Learning: Self-Evolving LLMs through Meta-Evaluation
by: Sui, Yuan, et al.
Published: (2026)
by: Sui, Yuan, et al.
Published: (2026)
Enhancing Reliability across Short and Long-Form QA via Reinforcement Learning
by: Wang, Yudong, et al.
Published: (2025)
by: Wang, Yudong, et al.
Published: (2025)
Breaking the Cycle of Recurring Failures: Applying Generative AI to Root Cause Analysis in Legacy Banking Systems
by: Jin, Siyuan, et al.
Published: (2024)
by: Jin, Siyuan, et al.
Published: (2024)
ShieldLM: Empowering LLMs as Aligned, Customizable and Explainable Safety Detectors
by: Zhang, Zhexin, et al.
Published: (2024)
by: Zhang, Zhexin, et al.
Published: (2024)
Why Do Large Language Models (LLMs) Struggle to Count Letters?
by: Fu, Tairan, et al.
Published: (2024)
by: Fu, Tairan, et al.
Published: (2024)
Similar Items
-
A Comparison of DeepSeek and Other LLMs
by: Gao, Tianchen, et al.
Published: (2025) -
Knowledge Graph-Driven Retrieval-Augmented Generation: Integrating Deepseek-R1 with Weaviate for Advanced Chatbot Applications
by: Lecu, Alexandru, et al.
Published: (2025) -
A Status Quo Investigation of Large Language Models towards Cost-Effective CFD Automation with OpenFOAMGPT: ChatGPT vs. Qwen vs. Deepseek
by: Wang, Wenkang, et al.
Published: (2025) -
Can generative AI figure out figurative language? The influence of idioms on essay scoring by ChatGPT, Gemini, and Deepseek
by: Oğuz, Enis
Published: (2025) -
A comprehensive study of LLM-based argument classification: from LLAMA through GPT-4o to Deepseek-R1
by: Pietroń, Marcin, et al.
Published: (2025)