In 6 months or less, when DeepSeek or Alibaba releases a <100B looped MoE model with <30B n-gram that matches GPT6-Astra and Fable 5.1, what then?
Article automatically generated from technical news.
With all the different optimizations and firepowers we've seen from open weights so far. What then? China only steals/distill from OpenAI and Anthropic? Well, clearly there's something wrong with AA? Benchmarks don't mean anything? They just benchmaxxed? Whatever, OpenAI/Anthropic has moved onto AAGI or AAAGI or AAAAGI or AAAAAGI? I can't run this locally on my toaster desktop anyways? submitted by /u/Trollsofalabama to r/LocalLLM [link] &
Fonte originale