reddit/r/machinelearningnews

Does llms.txt actually help a model read a site? We ran a sealed 3-condition bench on 8 arms (4 local Q4, 4 frontier cloud) — and published the number that cuts against our own registered result

u/strata2signal 2026-08-17

A research lab conducted a sealed benchmark testing whether llms.txt files improve model performance across 8 arms (4 local Q4 and 4 frontier cloud models) using 30 standardized questions about their own website content.…

→ View original source
reddit/r/machinelearningnews

A continuous dynamical system drives word embeddings to a stable equilibrium — 4 statistical invariants stay constant, and word relations get 4× stronger vs. controls

u/Ok_Department_4063 2026-07-17

Research demonstrates that pre-trained word embeddings evolve through a continuous dynamical system into a stable equilibrium, marked by four invariant statistical properties. Over 5,000 integration steps using GloVe 300…

→ View original source
reddit/r/machinelearningnews

Moonshot AI just released Kimi K3. It is a 2.8-trillion-parameter model with native vision and a 1-million-token context window. Moonshot calls it the world’s first open 3T-class model.

u/ai-lover 2026-07-17

Moonshot AI has launched Kimi K3, a 2.8‑trillion‑parameter open MoE model that it claims is the world’s first 3‑trillion‑parameter open model. The system incorporates Kimi Delta Attention, a hybrid linear attention mecha…

→ View original source