this post was submitted on 01 Dec 2024
97 points (81.3% liked)

Technology

60093 readers
2807 users here now

This is a most excellent place for technology news and articles.


Our Rules


  1. Follow the lemmy.world rules.
  2. Only tech related content.
  3. Be excellent to each another!
  4. Mod approved content bots can post up to 10 articles per day.
  5. Threads asking for personal tech support may be deleted.
  6. Politics threads may be removed.
  7. No memes allowed as posts, OK to post as comments.
  8. Only approved bots from the list below, to ask if your bot can be added please contact us.
  9. Check for duplicates before posting, duplicates may be removed

Approved Bots


founded 2 years ago
MODERATORS
 

cross-posted from: https://futurology.today/post/2910566

Alibaba's Qwen team just released QwQ-32B-Preview, a powerful new open-source AI reasoning model that can reason step-by-step through challenging problems and directly competes with OpenAI's o1 series across benchmarks.

The details:

QwQ features a 32K context window, outperforming o1-mini and competing with o1-preview on key math and reasoning benchmarks.

The model was tested across several of the most challenging math and programming benchmarks, showing major advances in deep reasoning.

QwQ demonstrates ‘deep introspection,’ talking through problems step-by-step and questioning and examining its own answers to reason to a solution.

The Qwen team noted several issues in the Preview model, including getting stuck in reasoning loops, struggling with common sense, and language mixing.

Why it matters: Between QwQ and DeepSeek, open-source reasoning models are here — and Chinese firms are absolutely cooking with new models that nearly match the current top closed leaders. Has OpenAI’s moat dried up, or does the AI leader have something special up its sleeve before the end of the year?

top 50 comments
sorted by: hot top controversial new old
[–] [email protected] 5 points 3 weeks ago (1 children)

Qwen is super powerful but the CCP endorsed lobodomy to censor it make it useless for my needs. Mistral 22B > qwen 32B all day any day just because it won't shriek at me in rejection when I ask the wrong question.

[–] [email protected] 2 points 3 weeks ago

Surprisingly asking about Jack Ma's resignation didn't stop it, it mentioned that he was allegedly forced to step down for political reasons. When I probed further on what these were it didn't error but did randomly switch to Chinese at one point in the answer. image

It's very verbose compared to any other AI I've used. Not necessarily all great but there's some good parts in the responses. I'm more curious how it was trained behind the great firewall.

[–] [email protected] 19 points 3 weeks ago (1 children)

It's not fair to describe the western models as "closed". All the tech bros have a open-source ethic that would embarrass a FOSS developer. At least when it comes to training data.

[–] [email protected] 17 points 3 weeks ago

Har har. "I just took it, you can have it too."

load more comments
view more: next ›