Sample Count Is Not Enough: Candidate-Generation Strategy Shapes the Energy and Performance of LLM Test-Time Scaling Paper • 2609.19499 • Published 5 days ago • 29 • 4
ZGCM-1: A Fully Open and Extremely Efficient Foundation Model for Math and Agentic Search Paper • 2609.13356 • Published 10 days ago • 256 • 8
Thinking in a Low-Resource Language: What SFT Builds, What RL Fixes, What Accuracy Cannot See Paper • 2608.17744 • Published Aug 18 • 16 • 3
Teaching Nemotron Greek: Mining a Corpus, Adapting Retrieval, and Grounding Generation for Modern Greek across Specialist Domains Paper • 2608.05138 • Published Aug 5 • 32 • 3
MameLoshnLM: Yiddish Language Model and Evaluation Benchmark Paper • 2608.05850 • Published Aug 6 • 23 • 3
MindForge: Teaching Small Language Models Whole-Life-Cycle Software Engineering via Source-Free Program Synthesis Paper • 2607.27146 • Published Jul 29 • 30 • 3
MobileLLM-R1: Exploring the Limits of Sub-Billion Language Model Reasoners with Open Training Recipes Paper • 2509.24945 • Published Sep 29, 2025 • 7 • 1
Smarter and Cheaper at Once: Byte-Exact KV-Cache Grafting Turns a Frozen Small Model into a Verified-Knowledge Flywheel Paper • 2607.14431 • Published Jul 15 • 7 • 7
MuScriptor: An Open Model for Multi-Instrument Music Transcription Paper • 2607.08168 • Published Jul 9 • 17 • 3
Phone Segmentation and Recognition through Phonological Activation Mapping Paper • 2607.09020 • Published Jul 10 • 5 • 3
Can Dialects Be Steered Like Languages? Sparse Neurons and Distributed Directions in Arabic LLMs Paper • 2607.03936 • Published Jul 4 • 3 • 3
UP: Unbounded Positive Asymmetric Optimization for Breaking the Exploration-Stability Dilemma Paper • 2607.06987 • Published Jul 8 • 9 • 3
Formalizing Latent Thoughts: Four Axioms of Thought Representation in LLMs Paper • 2606.27378 • Published May 7 • 61 • 7
Formalizing Latent Thoughts: Four Axioms of Thought Representation in LLMs Paper • 2606.27378 • Published May 7 • 61 • 7
Formalizing Latent Thoughts: Four Axioms of Thought Representation in LLMs Paper • 2606.27378 • Published May 7 • 61 • 7
Morpheus: A Morphology-Aware Neural Tokenizer and Word Embedder for Turkish Paper • 2606.18717 • Published Jun 17 • 5 • 3