Qwen/Alibaba have stopped doing open weights releases for a while. No grudge or anything, I'm certainly not going to look at a gift horse in the mouth, but both DeepSeek and Moonshot have been very consistent with open weights as well as sharing actually detailed research.
In terms of open research, China has absolutely overtaken the US.
Qwen has a pinned tweet stating that 3.8 will be released as open weights soon. I guess it remains to be seen, though, if they’ll do the smaller model sizes or only the big 2.4T one.
Could these complex/hard to read Fable outputs be sign of some kind of industrial level of intelligence, which us humans may have a hard to comprehend, while it may be also hard for machine to use simpler texts to properly outline all nuances and complexities of concepts it output?
(1) it’s not that I can’t understand their output, it’s just written in a way that is very homogenous and same-y, with very boring cliches and phrases that don’t quite match their context
(2) a pretty strong sign of intelligence is being able to explain complex things in simple terms
It's very much number 2 in my experience with it. Predominantly whenever a really advanced word or phrase is used it can be substituted for a much simpler one without losing useful context. And it seems to really prefer to use them a lot.
It's more like Claude models entirely suck at extracting key points. No matter how hard I emphasize that it needs to pick the "load-bearing" facts and claims, it cannot stop itself muttering around. It never nails the core logical structure. GPT is better at that.
If you feed Fable or Opus primarily handoff documents from a previous context instead of human written prompts and are working on something sophisticated it rapidly reaches a level where it's hard to actually comprehend for a non expert. I've received incredibly obtuse outputs that contain more mathematical formulas than English words with programming workloads.
reply