this post was submitted on 30 Aug 2026
26 points (88.2% liked)

LocalLLaMA

5120 readers
10 users here now

Welcome to LocalLLaMA! Here we discuss running and developing machine learning models at home. Lets explore cutting edge open source neural network technology together.

Get support from the community! Ask questions, share prompts, discuss benchmarks, get hyped at the latest and greatest model releases! Enjoy talking about our awesome hobby.

As ambassadors of the self-hosting machine learning community, we strive to support each other and share our enthusiasm in a positive constructive way.

Rules:

Rule 1 - No harassment or personal character attacks of community members. I.E no namecalling, no generalizing entire groups of people that make up our community, no baseless personal insults.

Rule 2 - No comparing artificial intelligence/machine learning models to cryptocurrency. I.E no comparing the usefulness of models to that of NFTs, no comparing the resource usage required to train a model is anything close to maintaining a blockchain/ mining for crypto, no implying its just a fad/bubble that will leave people with nothing of value when it burst.

Rule 3 - No comparing artificial intelligence/machine learning to simple text prediction algorithms. I.E statements such as "llms are basically just simple text predictions like what your phone keyboard autocorrect uses, and they're still using the same algorithms since <over 10 years ago>.

Rule 4 - No implying that models are devoid of purpose or potential for enriching peoples lives.

founded 3 years ago
MODERATORS
 

Qwen3.6-27B and Qwen3.8-27B are great. They are fast and quite accurate for not having to spend too much time thinking or searching, as other models with comparable scores have to.

But Qwen3.8-Flash-Next, which tops a lot of benchmarks in its category, is the most obvious Claude distill ever, and although I don't particularly care about how models get trained, I hate the way it sounds and interacts with me.

  • Three warts I'd still fix
  • Statements are down to the two things code can't say
  • The bug is fixed at the substrate that caused it
  • Genuinely good now
  • Also worth a conscious nod
  • Say the word on 1 and/or 3 and I'll do them

Not to mention Claude's over-the-top comments and git commit messages that make you want to scream.

I wish Alibaba would go back to adopting its own style. Qwen3.6-27B and Qwen3.8-27B sound much better.

you are viewing a single comment's thread
view the rest of the comments
[–] pepperfree@sh.itjust.works 1 points 3 days ago* (last edited 3 days ago)

I usually use skill like openspec when requesting LLM to do something, combining them with matt's skills like grill-me. I assume both of this skill will make LLM output in certain way as I jump between model, I don't really notice them. Usually I also ask to 'talk in basic english language' or ' talk in simplified technical english language' when just asking random things in the codebase. Surely not perfect and better to have normal sounding LLM directly.