Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Paper β’ 2607.27919 β’ Published 4 days ago β’ 49
Every Time I Hire a Linguist, Inference Costs Go Down: On Linguistic Rules as Effective Prompt Compressors Paper β’ 2607.25335 β’ Published 6 days ago β’ 1
From Data to Device: ELMOD An Efficient German-First 2.7B Language Model for Mobile Inference Paper β’ 2607.24585 β’ Published 7 days ago β’ 1
ELMOD-2.7b Collection Efficient Language Model for On-Device Deployment (ELMOD) a compact (2.7B) German LM designed for efficient inference on constrained hardware. β’ 2 items β’ Updated 7 days ago β’ 2
Building a European Multilingual Evaluation Dataset: The MMLU Localisation Project within the EMT Network Paper β’ 2607.18432 β’ Published 14 days ago β’ 2
Separating Representation from Reconstruction Enables Scalable Text Encoders Paper β’ 2607.04011 β’ Published about 1 month ago β’ 1
Apertus v1.5 Collection Democratizing Open and Compliant LLMs for Global Language Environments: 8B and 70B open-data open-weights models, multilingual in >1000 languages β’ 2 items β’ Updated about 23 hours ago β’ 27
news-crawler-LM: A Small Long-Context Model For High-Quality News Crawling Paper β’ 2607.21284 β’ Published 11 days ago β’ 2
view article Article Be Ready Before the Attack: A Practical Guide to Self-Hosting an Open Model for Cyber Defense jeffboudier β’ 13 days ago β’ 20
A Sovereign, Open-Source Foundation Model for German and English Paper β’ 2607.09424 β’ Published 24 days ago β’ 15
view article Article Welcome Inkling by Thinking Machines +3 burtenshaw, merve, pcuenq, ariG23498, andito β’ 19 days ago β’ 145
view article Article Native-speed vLLM transformers modeling backend hmellor, lysandre β’ 26 days ago β’ 64
KVpop -- Key-Value Cache Compression with Predictive Online Pruning Paper β’ 2607.05061 β’ Published 28 days ago β’ 24
view article Article Hugging Face and Cerebras bring Gemma 4 to real-time voice AI +2 A-Mahla, andito, lvwerra, vyassaurabh β’ Jul 1 β’ 90
Apertus Mini Collection Distillations and Quantizations of our models into more compact formats (<8B parameters) β’ 17 items β’ Updated Jun 24 β’ 12
Do We Still Need Fine Tuning? Turkish Sentiment Analysis in the Era of Large Language Model Paper β’ 2606.29614 β’ Published Jun 28 β’ 1