Why your local LLM feels dumber than it is (forum.level1techs.com)

25 points by felineflock 2 hours ago

jonplackett 26 minutes ago

I just got qwen 3.8 27b mlx running on my Macbook Pro and honestly I’m pretty blown away by how not-dumb it is.

alexchantavy 4 minutes ago

How many tok/s are you getting? What gen mbp?

prettyblocks 12 minutes ago

My problem is how hot they run. I'm an m4 pro. Do you have the same issue?

lukan 6 minutes ago

I don't have the hardware but a often mentioned advice is to put your mac into energy saving mode - it still will work, a bit slower, but stays cool.

downrightmike 11 minutes ago

Mineral oil bath?

StarlaAtNight 11 minutes ago

how quick does it respond? what are specs of your laptop?

chorlton2080 4 minutes ago

Does it need to respond fast? For important applications, I'm sure we'd all be fine waiting 20 minutes for a high quality, usable answer. Or is it the need for interative refinements that make speed relevant?