I usually use skill like openspec when requesting LLM to do something, combining them with matt's skills like grill-me. I assume both of this skill will make LLM output in certain way as I jump between model, I don't really notice them. Usually I also ask to 'talk in basic english language' or ' talk in simplified technical english language' when just asking random things in the codebase. Surely not perfect and better to have normal sounding LLM directly.
LocalLLaMA
Welcome to LocalLLaMA! Here we discuss running and developing machine learning models at home. Lets explore cutting edge open source neural network technology together.
Get support from the community! Ask questions, share prompts, discuss benchmarks, get hyped at the latest and greatest model releases! Enjoy talking about our awesome hobby.
As ambassadors of the self-hosting machine learning community, we strive to support each other and share our enthusiasm in a positive constructive way.
Rules:
Rule 1 - No harassment or personal character attacks of community members. I.E no namecalling, no generalizing entire groups of people that make up our community, no baseless personal insults.
Rule 2 - No comparing artificial intelligence/machine learning models to cryptocurrency. I.E no comparing the usefulness of models to that of NFTs, no comparing the resource usage required to train a model is anything close to maintaining a blockchain/ mining for crypto, no implying its just a fad/bubble that will leave people with nothing of value when it burst.
Rule 3 - No comparing artificial intelligence/machine learning to simple text prediction algorithms. I.E statements such as "llms are basically just simple text predictions like what your phone keyboard autocorrect uses, and they're still using the same algorithms since <over 10 years ago>.
Rule 4 - No implying that models are devoid of purpose or potential for enriching peoples lives.
I have to use Claude at work and have found myself wishing I could use Qwen 3.8 27B at work so I don’t have to waste time deciphering a machine’s fart sniffing. Guess it’s a good thing I don’t have enough VRAM to run a 125B model.
I haven't used claude in ages but I recall 4.5 and 4.6 haiku and sonnet being pretty great, tonally.
I'm guessing you mean 4.8 series / Fable?
Yes, Opus/Fable 5. If you have the opportunity, play around with it. It writes in a format that tries so hard to be concise (while blurting out a monumental essay) that you can hardly understand it, and repeats the same phrases all the time.
I hate Claude's voice so much. So much ceremony spent on a few simple questions and caveats. I feel like AI defaults to assuming you have no idea what you're doing. ChatGPT is the same but with different verbal tics.
Let's cut down to the heart of the matter as you hit the nail squarely on the head
Never used Claude, but GPT-5.6 Sol annoyed me so much I actively avoid it and just use Terra on max. Sol is trained to favor convoluted San Francisco jargon that projects any trivial idea into some sort of an ingenious business maneuver.
I only use Qwen for Zoo Code, when chatting I just use Gemma
That seems to be the consensus on Reddit, Qwen for code, Gemma for everything else. At least it was so for 3.6, not sure if 3.8 changes the situation.