<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0"><channel><title>Open Weight Intelligence</title><link>https://openmodelweights.com/news/</link><description>Daily source-first intelligence on open-weight AI.</description><language>en</language><item><title>Ai2 releases AstaBrief 8B for fast, cited scientific report generation</title><link>https://openmodelweights.com/news/ai2-astabrief-8b-open-weight-scientific-reports/</link><guid isPermaLink="true">https://openmodelweights.com/news/ai2-astabrief-8b-open-weight-scientific-reports/</guid><pubDate>Fri, 02 Oct 2026 15:00:00 +0000</pubDate><category>Model release</category><description>AstaBrief turns retrieved scientific evidence into cited reports and gives institutions an open-weight option for running the report-writing stage on their own infrastructure.</description><source url="https://huggingface.co/blog/allenai/astabrief">Ai2 / Hugging Face</source></item><item><title>llama.cpp adds decision-model inference through a new System One endpoint</title><link>https://openmodelweights.com/news/llama-cpp-adds-decision-models-system-one-endpoint/</link><guid isPermaLink="true">https://openmodelweights.com/news/llama-cpp-adds-decision-models-system-one-endpoint/</guid><pubDate>Fri, 02 Oct 2026 14:30:00 +0000</pubDate><category>Runtime</category><description>The local inference stack can now score typed options directly instead of generating free-form text, opening a different deployment path for routing, moderation and agent control.</description><source url="https://huggingface.co/blog/ggml-org/decision-models-in-llamacpp">ggml-org / Hugging Face</source></item><item><title>Olmo-core 3 opens a training stack designed for trillion-scale mixture-of-experts models</title><link>https://openmodelweights.com/news/olmo-core-3-open-moe-training-stack/</link><guid isPermaLink="true">https://openmodelweights.com/news/olmo-core-3-open-moe-training-stack/</guid><pubDate>Fri, 02 Oct 2026 13:40:00 +0000</pubDate><category>Training infrastructure</category><description>Ai2 redesigned its open training infrastructure around MoE routing, expert parallelism and lower-precision execution, with the next generation of Olmo expected to use the stack.</description><source url="https://huggingface.co/blog/allenai/olmocore3">Ai2 / Hugging Face</source></item><item><title>NVIDIA Kumo Tabular brings open foundation models to tabular prediction</title><link>https://openmodelweights.com/news/nvidia-kumo-tabular-open-foundation-model/</link><guid isPermaLink="true">https://openmodelweights.com/news/nvidia-kumo-tabular-open-foundation-model/</guid><pubDate>Fri, 02 Oct 2026 12:50:00 +0000</pubDate><category>Model release</category><description>The Kumo Structured family targets classification and regression directly from labeled tables, with three small model sizes and a commercial-use license.</description><source url="https://huggingface.co/blog/nvidia/kumo-tabular">NVIDIA / Hugging Face</source></item><item><title>Holo4 ships 27B and 35B-A3B agent models with open weights and trajectories</title><link>https://openmodelweights.com/news/holo4-open-agent-models-27b-35b/</link><guid isPermaLink="true">https://openmodelweights.com/news/holo4-open-agent-models-27b-35b/</guid><pubDate>Fri, 02 Oct 2026 12:10:00 +0000</pubDate><category>Model release</category><description>H Company’s new agentic family is designed to move between GUIs, code, MCP and APIs, with downloadable weights in multiple precisions and public benchmark trajectories.</description><source url="https://huggingface.co/blog/Hcompany/holo4">H Company / Hugging Face</source></item><item><title>Liquid AI releases a 280M draft model to accelerate LFM2.5-VL-3B</title><link>https://openmodelweights.com/news/liquid-ai-lfm25-vl-dspark-speculative-decoding/</link><guid isPermaLink="true">https://openmodelweights.com/news/liquid-ai-lfm25-vl-dspark-speculative-decoding/</guid><pubDate>Fri, 02 Oct 2026 11:30:00 +0000</pubDate><category>Inference</category><description>The DSpark companion model adds speculative decoding to Liquid AI’s vision-language stack with integrations for llama.cpp, MLX-VLM and SGLang.</description><source url="https://huggingface.co/blog/LiquidAI/lfm2-5-vl-dspark">Liquid AI / Hugging Face</source></item><item><title>Black Forest Labs releases FLUX 3 Action, a 7B open-weight world action model</title><link>https://openmodelweights.com/news/flux-3-action-open-weight-world-action-model/</link><guid isPermaLink="true">https://openmodelweights.com/news/flux-3-action-open-weight-world-action-model/</guid><pubDate>Fri, 02 Oct 2026 10:55:00 +0000</pubDate><category>Model release</category><description>FLUX 3 Action predicts future frames and actions together and arrives with robot-policy checkpoints, LeRobot integration and a dedicated weights license.</description><source url="https://huggingface.co/blog/black-forest-labs/flux-3-action">Black Forest Labs / Hugging Face</source></item><item><title>NVIDIA releases a 100M open-weight Nemotron 3 model for speaker diarization</title><link>https://openmodelweights.com/news/nvidia-nemotron-3-diarization-open-weight-100m/</link><guid isPermaLink="true">https://openmodelweights.com/news/nvidia-nemotron-3-diarization-open-weight-100m/</guid><pubDate>Fri, 02 Oct 2026 10:20:00 +0000</pubDate><category>Model release</category><description>Nemotron 3 Diarization targets real-time and offline speaker attribution, turning speech streams into speaker-aware timelines for meetings, calls and voice agents.</description><source url="https://huggingface.co/blog/nvidia/nemotron-diarization">NVIDIA / Hugging Face</source></item><item><title>Hugging Face Transformers adds efficient GGUF inference using llama.cpp kernels</title><link>https://openmodelweights.com/news/transformers-runs-gguf-llama-cpp-quants/</link><guid isPermaLink="true">https://openmodelweights.com/news/transformers-runs-gguf-llama-cpp-quants/</guid><pubDate>Fri, 02 Oct 2026 09:40:00 +0000</pubDate><category>Runtime</category><description>GGUF checkpoints can now be loaded through familiar Transformers APIs while reusing ggml kernels to target practical local inference, initially with an Apple Silicon focus.</description><source url="https://huggingface.co/blog/transformers-llama-cpp-quants">Hugging Face</source></item><item><title>NVIDIA agrees to acquire Hugging Face and says the platform will remain open</title><link>https://openmodelweights.com/news/nvidia-agrees-to-acquire-hugging-face-open-platform-pledge/</link><guid isPermaLink="true">https://openmodelweights.com/news/nvidia-agrees-to-acquire-hugging-face-open-platform-pledge/</guid><pubDate>Fri, 02 Oct 2026 09:00:00 +0000</pubDate><category>Ecosystem</category><description>The $12.93 billion agreement would put the central model-sharing platform inside NVIDIA while the companies publicly commit to model, cloud and accelerator choice.</description><source url="https://blogs.nvidia.com/blog/nvidia-to-acquire-hugging-face/">NVIDIA</source></item></channel></rss>
