NeoMME: A Single-Tower Multimodal-Native Multilingual Foundation Encoder for Efficient Fine-Tuning and Inference Paper β’ 2609.01657 β’ Published Aug 31 β’ 36
view article Article Lattice: an 8 MB static retriever that embeds Wikipedia in 7 minutes erikkaum β’ Aug 7 β’ 19
view article Article LFM2.5-Encoders for Fast Long-Context Inference on CPU LiquidAI β’ Jul 28 β’ 69
AMALIA LLM 0626 Collection AMALIA models and datasets released in June 2026 β’ 5 items β’ Updated 25 days ago β’ 8
view article Article Optimum Intel 2.0: An OpenVINO-First Toolkit for Running Open Models on Intel jeffboudier β’ Jun 11 β’ 5
view article Article Designing the hf CLI as an agent-optimized way to work with the Hub celinah, Wauplin β’ Jun 4 β’ 61
view article Article How to Fine-Tune Nemotron 3.5 ASR for Your Language, Domain, or Accent nvidia β’ Jun 4 β’ 79
nvidia/nemotron-3.5-asr-streaming-0.6b Automatic Speech Recognition β’ 0.6B β’ Updated 26 days ago β’ 1.28M β’ β’ 1.17k