Gecko4072 1 hour ago

DeepSwe score is 42.2. For comparison 3.6-27b is 13.3, GLM 5.2 is 44, and Opus 4.8 is 59.

bravetraveler 6 hours ago

Didn't think we'd have model launch pages so 'soon'; how long until we have something Steam-like and can preload our eagerly-awaited models?

Anyway, eager to try [on hardware I own]!

WithinReason 1 hour ago

Trading blows with Opus 4.6 is impressive, awaiting unsloth quants

meffmadd 6 hours ago

Hopefully they release the 35B-A3B version alongside it.

nezhar 3 hours ago

It feels so strange that we now have countdowns for model releases

  • jimmydoe 2 hours ago

    Alibaba does this for their shopping biz all the time. Growth hack is very high priority in the company. Their homepage was a bunch of ridiculous SEO. The whole culture there seems very distasteful to me. This doesn’t mean this qwen model is bad today, but I’m a believer that culture will decide the product quality in the long run.

wolvoleo 9 hours ago

I miss a good 9B class model :',( The last one was 3.5.

27B is just a little bit too big for a 16GB GPU.

  • nezhar 3 hours ago

    Yeah, same with the latest releases of meta and NVIDIA. It's like everybody is expected to have 32 GB of RAM

darthoctopus 10 hours ago

this page gives a 404

  • kristjansson 10 hours ago

    Oh shoot it had a nice countdown timer to ~8a PT when I submitted it