reddit/r/LocalLLaMA

MiniCPM5-2B Release Day

u/Equivalent-Grass-527 2026-09-07

OpenBMB has released MiniCPM5-2B, a new open-weights language model that achieves a score of 15 on the Artificial Analysis Intelligence Index v4.2, marking the highest score among open-weight models at 4B parameters or b…

→ View original source
reddit/r/LocalLLaMA

New Gemma models on arena ai

u/Hot_Example_4456 2026-09-01

The Reddit post announces the release of new Gemma models on the Arena AI platform, noting uncertainty about whether the update includes Gemma 5 or other variants. The brief comment includes an attached image and is dire…

→ View original source
reddit/r/LocalLLaMA

Copilot you say?

u/edge_compute_user 2026-08-24

The post titled "Copilot you say?" on r/LocalLLaMA describes a method for talking to any white‑collar employee using a LocalLLaMA model. It was submitted by user u/edge_compute_user on 2026‑08‑24. Read original

→ View original source
reddit/r/LocalLLaMA

Qwen 3.8 distillations

u/jacek2023 2026-08-16

Here's a thinking process: 1. **Analyze the Request:** - Role: Technical news summarizer - Task: Condense provided news into brief HTML summary - Input: Title, Source, URL, Author, Date, Description/Content - Output form…

→ View original source
reddit/r/LocalLLaMA

Is Microsoft-Phi dead?

u/Dance-Till-Night1 2026-08-08

The Microsoft‑Phi family, once a popular small‑model line, last saw a major release in December 2024, followed by a series of Phi 4 iterations such as the Phi 4‑reasoning‑vision launched in March 2026. Community members …

→ View original source
reddit/r/LocalLLaMA

LFM2.5-2.6B is out

u/Alarming_Positive_59 2026-08-04

Released today, with emphasis on agentic capabilities. I really like their models for simple, high volume tasks ("summarize these gazillion documents") and their 8b-a1b was my go-to for c

→ View original source
reddit/r/LocalLLaMA

MiniMax-H3 now on huggingface

u/Mobile-Pumpkin7944 2026-08-03

MiniMax H3 is an omni-modal generative system capable of understanding and generating content across text, images, video, and audio. It generates video with native stereo audio at resolutions up to 2K and durations of up…

→ View original source
reddit/r/LocalLLaMA

Qwen 3.8 is live now.

u/Mobile-Pumpkin7944 2026-08-03

Qwen has released Qwen 3.8, a 2.4-trillion-parameter Mixture-of-Experts (MoE) model designed for advanced coding and professional tasks. The model is available at https://www.qwencloud.com/models/qwen3.8-max, with open w…

→ View original source
reddit/r/LocalLLaMA

Jensen Huang: During the Hugging Face incident, closed AI blocked essential forensics. An open-weight frontier model helped contain the intrusion. That’s why we created the Open Secure AI Alliance.

u/Nunki08 2026-07-27

Jensen Huang said that during the Hugging Face incident, closed AI blocked essential forensics. An open-weight frontier model helped contain the intrusion, which led to the creation of the Open Secure AI Alliance. Read o…

→ View original source
reddit/r/LocalLLaMA

AMD Instella-MoE-16B-A3B

u/Look_0ver_There 2026-07-24

AMD has released a new model named Instella-MoE-16B-A3B, currently hosted on HuggingFace. This development marks AMD's entry into the open-source large language model space. The model utilizes a Mixture-of-Experts (MoE) …

→ View original source
reddit/r/LocalLLaMA

Kimi K3 just fixed 15 critical security bugs that Codex and Fable refused because of “cyber guardrails”. Hugging Face: We had this experience ourselves this week! Very scary to be guardrailed as a defender when you know attackers are likely bypassing

u/Nunki08 2026-07-20

David Sacks on 𝕏: https://x.com/DavidSacks/status/2078984980588531855 calle on 𝕏: https://x.com/callebtc/status/2078574362316165611 clem 🤗 on 𝕏: https://x.com/ClementDelangue

→ View original source