Model Announcements & Updates
Articles covering major model releases, announcement dates, feature upgrades, pricing changes, and product updates for LLMs from Anthropic, OpenAI, Google, Meta, and others.
Fine-Tuning Small LLMs: What the Research Really Says
Verified findings from arXiv:2412.13337 on optimal supervised fine-tuning strategies for 3B–7B parameter models
Meta Trains Code Llama to Reason Like a Compiler
Meta's LLM Compiler builds on Code Llama with 546B tokens of LLVM IR and assembly training to target code size optimization and compiler reasoning.
Community-Made LLM Chat Streaming Showcase on Hugging Face
A showcase of olivierdehaene's Hugging Face Space, where the community can discover and interact with multiple large language models through a shared streaming interface.
The 65B Fine-Tuned Champion and the 7B Base Merge King
A curated overview of the highest-performing fine-tuned (🔶) and base merges/moerges (🤝) models approximately 65B and 7B parameters, respectively, as ranked on the Hugging Face Open LLM Leaderboard collection as of March 2025.
Greenie Sweden's Grammar-Focused LLM: Open Source AI Advancement
A fine-tuned 1B parameter grammar-correction LLM by greenie-sweden, built on Unsloth and HuggingFace TRL, released under Apache 2.0 with deployment guides for Transformers, llama.cpp, Docker, and more.
Alibaba's Qwen3.8-Max Targets Long-Horizon Enterprise Automation With 2.4T Parameters
Alibaba's 2.4T Qwen3.8-Max model leads on agentic benchmarks, offers $8/M token pricing, and promises open weights — but independent verification is pending.
Liquid AI's LFM2.5-2.6B: Small Models, Big Edge Promises
Liquid AI released LFM2.5-2.6B, a 2.6 billion parameter open-weight model optimized for edge device deployment. Running on CPUs from Raspberry Pi to Apple Silicon without cloud dependency, the model targets agentic workloads with native tool calling, a 128K context window, and a revenue-gated open license.