bravetraveler 1 day ago

Didn't think we'd have model launch pages so 'soon'; how long until we have something Steam-like and can preload our eagerly-awaited models?

Anyway, eager to try [on hardware I own]!

Gecko4072 20 hours ago

DeepSwe score is 42.2. For comparison 3.6-27b is 13.3, GLM 5.2 is 44, and Opus 4.8 is 59.

meffmadd 1 day ago

Hopefully they release the 35B-A3B version alongside it.

wolvoleo 1 day ago

I miss a good 9B class model :',( The last one was 3.5.

27B is just a little bit too big for a 16GB GPU.

  • nezhar 21 hours ago

    Yeah, same with the latest releases of meta and NVIDIA. It's like everybody is expected to have 32 GB of RAM

WithinReason 20 hours ago

Trading blows with Opus 4.6 is impressive, awaiting unsloth quants

nezhar 21 hours ago

It feels so strange that we now have countdowns for model releases

  • jimmydoe 21 hours ago

    Alibaba does this for their shopping biz all the time. Growth hack is very high priority in the company. Their homepage was a bunch of ridiculous SEO. The whole culture there seems very distasteful to me. This doesn’t mean this qwen model is bad today, but I’m a believer that culture will decide the product quality in the long run.