-
Large Language Models are Locally Linear Mappings
Paper • 2505.24293 • Published • 15 -
Vision-Language-Vision Auto-Encoder: Scalable Knowledge Distillation from Diffusion Models
Paper • 2507.07104 • Published • 45 -
KV Cache Steering for Inducing Reasoning in Small Language Models
Paper • 2507.08799 • Published • 40
Collections
Discover the best community collections!
Collections including paper arxiv:2505.24293
-
Embodied Agents Meet Personalization: Exploring Memory Utilization for Personalized Assistance
Paper • 2505.16348 • Published • 53 -
Exploring the Latent Capacity of LLMs for One-Step Text Generation
Paper • 2505.21189 • Published • 62 -
Large Language Models are Locally Linear Mappings
Paper • 2505.24293 • Published • 15 -
Can LLMs Identify Critical Limitations within Scientific Research? A Systematic Evaluation on AI Research Papers
Paper • 2507.02694 • Published • 18
-
CoRAG: Collaborative Retrieval-Augmented Generation
Paper • 2504.01883 • Published • 10 -
SQL-R1: Training Natural Language to SQL Reasoning Model By Reinforcement Learning
Paper • 2504.08600 • Published • 30 -
Reasoning-SQL: Reinforcement Learning with SQL Tailored Partial Rewards for Reasoning-Enhanced Text-to-SQL
Paper • 2503.23157 • Published • 11 -
AI Agents: Evolution, Architecture, and Real-World Applications
Paper • 2503.12687 • Published • 2
-
microsoft/bitnet-b1.58-2B-4T
Text Generation • 0.8B • Updated • 5.88k • 1.16k -
M1: Towards Scalable Test-Time Compute with Mamba Reasoning Models
Paper • 2504.10449 • Published • 14 -
nvidia/Llama-3.1-Nemotron-8B-UltraLong-2M-Instruct
Text Generation • 8B • Updated • 1.15k • 15 -
ReTool: Reinforcement Learning for Strategic Tool Use in LLMs
Paper • 2504.11536 • Published • 61
-
I Have Covered All the Bases Here: Interpreting Reasoning Features in Large Language Models via Sparse Autoencoders
Paper • 2503.18878 • Published • 121 -
Large Language Models are Locally Linear Mappings
Paper • 2505.24293 • Published • 15 -
Thought Anchors: Which LLM Reasoning Steps Matter?
Paper • 2506.19143 • Published • 12
-
Large Language Models are Locally Linear Mappings
Paper • 2505.24293 • Published • 15 -
Vision-Language-Vision Auto-Encoder: Scalable Knowledge Distillation from Diffusion Models
Paper • 2507.07104 • Published • 45 -
KV Cache Steering for Inducing Reasoning in Small Language Models
Paper • 2507.08799 • Published • 40
-
Embodied Agents Meet Personalization: Exploring Memory Utilization for Personalized Assistance
Paper • 2505.16348 • Published • 53 -
Exploring the Latent Capacity of LLMs for One-Step Text Generation
Paper • 2505.21189 • Published • 62 -
Large Language Models are Locally Linear Mappings
Paper • 2505.24293 • Published • 15 -
Can LLMs Identify Critical Limitations within Scientific Research? A Systematic Evaluation on AI Research Papers
Paper • 2507.02694 • Published • 18
-
microsoft/bitnet-b1.58-2B-4T
Text Generation • 0.8B • Updated • 5.88k • 1.16k -
M1: Towards Scalable Test-Time Compute with Mamba Reasoning Models
Paper • 2504.10449 • Published • 14 -
nvidia/Llama-3.1-Nemotron-8B-UltraLong-2M-Instruct
Text Generation • 8B • Updated • 1.15k • 15 -
ReTool: Reinforcement Learning for Strategic Tool Use in LLMs
Paper • 2504.11536 • Published • 61
-
CoRAG: Collaborative Retrieval-Augmented Generation
Paper • 2504.01883 • Published • 10 -
SQL-R1: Training Natural Language to SQL Reasoning Model By Reinforcement Learning
Paper • 2504.08600 • Published • 30 -
Reasoning-SQL: Reinforcement Learning with SQL Tailored Partial Rewards for Reasoning-Enhanced Text-to-SQL
Paper • 2503.23157 • Published • 11 -
AI Agents: Evolution, Architecture, and Real-World Applications
Paper • 2503.12687 • Published • 2
-
I Have Covered All the Bases Here: Interpreting Reasoning Features in Large Language Models via Sparse Autoencoders
Paper • 2503.18878 • Published • 121 -
Large Language Models are Locally Linear Mappings
Paper • 2505.24293 • Published • 15 -
Thought Anchors: Which LLM Reasoning Steps Matter?
Paper • 2506.19143 • Published • 12