{"exhaustive":{"nbHits":false,"typo":false},"exhaustiveNbHits":false,"exhaustiveTypo":false,"hits":[{"_highlightResult":{"author":{"matchLevel":"none","matchedWords":[],"value":"somenameforme"},"comment_text":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["unsloth"],"value":"No idea where you're getting this from. The paleolithic was during an ice age full of big fatty game like <em>sloth</em>s, mammoths, and so on. Some believe we hunted these animals to the point of extinction with a <i>tiny</i> population, which would imply a very high level, arguably unbelievably high, of consumption of meat. Even if we didn't hunt them to extinction we certainly did engage in relatively large scale hunting and consumption of vast quantities of them, which already largely falsifies your concept.<p>Meat only became more of a luxury in the agricultural age where we started squeezing more and more people into smaller and smaller regions which meant there was no longer enough to go around. So plebs got the carbs, and royalty got the protein. FWIW you can also see this in the groups that have ancient traditions dating back far into the past, like the various tribal groups in Russia, Mongolian nomads, etc. Meat, and many highly creative ways of preserving it [1], remain the majority of their diet. They probably don't date back to the paleo, but at least far back enough to falsify the claim that animal fats are even remotely novel.<p>[1] - <a href=\"https://www.rbth.com/russian-kitchen/334150-dish-kill-kopalkhen-north\" rel=\"nofollow\">https://www.rbth.com/russian-kitchen/334150-dish-kill-kopalk...</a>"},"story_title":{"matchLevel":"none","matchedWords":[],"value":"Semaglutide linked to lower predicted dementia risk"},"story_url":{"matchLevel":"none","matchedWords":[],"value":"https://alz-journals.onlinelibrary.wiley.com/doi/10.1002/dad2.70432"}},"_tags":["comment","author_somenameforme","story_49311651"],"author":"somenameforme","comment_text":"No idea where you&#x27;re getting this from. The paleolithic was during an ice age full of big fatty game like sloths, mammoths, and so on. Some believe we hunted these animals to the point of extinction with a <i>tiny</i> population, which would imply a very high level, arguably unbelievably high, of consumption of meat. Even if we didn&#x27;t hunt them to extinction we certainly did engage in relatively large scale hunting and consumption of vast quantities of them, which already largely falsifies your concept.<p>Meat only became more of a luxury in the agricultural age where we started squeezing more and more people into smaller and smaller regions which meant there was no longer enough to go around. So plebs got the carbs, and royalty got the protein. FWIW you can also see this in the groups that have ancient traditions dating back far into the past, like the various tribal groups in Russia, Mongolian nomads, etc. Meat, and many highly creative ways of preserving it [1], remain the majority of their diet. They probably don&#x27;t date back to the paleo, but at least far back enough to falsify the claim that animal fats are even remotely novel.<p>[1] - <a href=\"https:&#x2F;&#x2F;www.rbth.com&#x2F;russian-kitchen&#x2F;334150-dish-kill-kopalkhen-north\" rel=\"nofollow\">https:&#x2F;&#x2F;www.rbth.com&#x2F;russian-kitchen&#x2F;334150-dish-kill-kopalk...</a>","created_at":"2026-08-16T06:36:05Z","created_at_i":1786862165,"objectID":"49317455","parent_id":49317221,"story_id":49311651,"story_title":"Semaglutide linked to lower predicted dementia risk","story_url":"https://alz-journals.onlinelibrary.wiley.com/doi/10.1002/dad2.70432","updated_at":"2026-08-16T11:16:12Z"},{"_highlightResult":{"author":{"matchLevel":"none","matchedWords":[],"value":"oblio"},"comment_text":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["unsloth"],"value":"To add to that. Those magical 50s and 60s?<p>Basically the entire industrial base in Eurasia had been destroyed during WW2. Unsurprisingly the only <em>unscath</em>ed industrialized country in the world entered an unprecedented boom. Assuming that everyone else would revert to the Stone Age forever was unrealistic."},"story_title":{"matchLevel":"none","matchedWords":[],"value":"First human trials of designer protein therapies stun US neuroscientists"},"story_url":{"matchLevel":"none","matchedWords":[],"value":"https://cen.acs.org/biological-chemistry/biotechnology/human-trial-chemogenetic-brain-therapy/104/web/2026/08"}},"_tags":["comment","author_oblio","story_49313097"],"author":"oblio","children":[49314082],"comment_text":"To add to that. Those magical 50s and 60s?<p>Basically the entire industrial base in Eurasia had been destroyed during WW2. Unsurprisingly the only unscathed industrialized country in the world entered an unprecedented boom. Assuming that everyone else would revert to the Stone Age forever was unrealistic.","created_at":"2026-08-15T20:28:31Z","created_at_i":1786825711,"objectID":"49314022","parent_id":49313914,"story_id":49313097,"story_title":"First human trials of designer protein therapies stun US neuroscientists","story_url":"https://cen.acs.org/biological-chemistry/biotechnology/human-trial-chemogenetic-brain-therapy/104/web/2026/08","updated_at":"2026-08-16T06:55:12Z"},{"_highlightResult":{"author":{"matchLevel":"none","matchedWords":[],"value":"me_bx"},"comment_text":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["unsloth"],"value":"Thank you and sibling poster for the suggestions.<p>I used the settings recommended on huggingface/<em>unsloth</em>'s model page + whatever suggestion from various LLMs - didn't research too much myself which setting did what."},"story_title":{"matchLevel":"none","matchedWords":[],"value":"Qwen 3.8 27B"},"story_url":{"matchLevel":"none","matchedWords":[],"value":"https://huggingface.co/Qwen/Qwen3.8-27B-FP8"}},"_tags":["comment","author_me_bx","story_49299605"],"author":"me_bx","comment_text":"Thank you and sibling poster for the suggestions.<p>I used the settings recommended on huggingface&#x2F;unsloth&#x27;s model page + whatever suggestion from various LLMs - didn&#x27;t research too much myself which setting did what.","created_at":"2026-08-15T20:24:57Z","created_at_i":1786825497,"objectID":"49313995","parent_id":49305771,"story_id":49299605,"story_title":"Qwen 3.8 27B","story_url":"https://huggingface.co/Qwen/Qwen3.8-27B-FP8","updated_at":"2026-08-15T20:27:56Z"},{"_highlightResult":{"author":{"matchLevel":"none","matchedWords":[],"value":"ChuckMcM"},"comment_text":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["unsloth"],"value":"Which sort of affirms the mantra, &quot;If you can't prove causation, try to prove correlation.&quot; This because you can market correlation to <em>unsoph</em>isticated readers and imply causation."},"story_title":{"matchLevel":"none","matchedWords":[],"value":"Semaglutide linked to 26% lower 5-year predicted dementia risk"},"story_url":{"matchLevel":"none","matchedWords":[],"value":"https://alz-journals.onlinelibrary.wiley.com/doi/10.1002/dad2.70432"}},"_tags":["comment","author_ChuckMcM","story_49311651"],"author":"ChuckMcM","children":[49314112,49313276],"comment_text":"Which sort of affirms the mantra, &quot;If you can&#x27;t prove causation, try to prove correlation.&quot; This because you can market correlation to unsophisticated readers and imply causation.","created_at":"2026-08-15T18:35:44Z","created_at_i":1786818944,"objectID":"49313083","parent_id":49312664,"story_id":49311651,"story_title":"Semaglutide linked to 26% lower 5-year predicted dementia risk","story_url":"https://alz-journals.onlinelibrary.wiley.com/doi/10.1002/dad2.70432","updated_at":"2026-08-15T20:38:40Z"},{"_highlightResult":{"author":{"matchLevel":"none","matchedWords":[],"value":"ssvegeta"},"comment_text":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["unsloth"],"value":"I'm seeing this infinite loop behavior during reasoning as well, for the exact same quant from <em>Unsloth</em>. Set presence penalty to 1 and the issue seemly went away, but having a penalty that high worries me for coding tasks. Will try the bartowski one now."},"story_title":{"matchLevel":"none","matchedWords":[],"value":"Qwen 3.8 27B"},"story_url":{"matchLevel":"none","matchedWords":[],"value":"https://huggingface.co/Qwen/Qwen3.8-27B-FP8"}},"_tags":["comment","author_ssvegeta","story_49299605"],"author":"ssvegeta","comment_text":"I&#x27;m seeing this infinite loop behavior during reasoning as well, for the exact same quant from Unsloth. Set presence penalty to 1 and the issue seemly went away, but having a penalty that high worries me for coding tasks. Will try the bartowski one now.","created_at":"2026-08-15T17:28:06Z","created_at_i":1786814886,"objectID":"49312449","parent_id":49304627,"story_id":49299605,"story_title":"Qwen 3.8 27B","story_url":"https://huggingface.co/Qwen/Qwen3.8-27B-FP8","updated_at":"2026-08-15T17:29:40Z"},{"_highlightResult":{"author":{"matchLevel":"none","matchedWords":[],"value":"dofm"},"comment_text":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["unsloth"],"value":"Have you tried turning down the new Qwen's reasoning effort level from xhigh, which it defaults at?<p>LM Studio isn't exposing a dropdown for this, at least with the <em>unsloth</em> build.<p><em>Unsloth</em> Studio / Desktop does."},"story_title":{"matchLevel":"none","matchedWords":[],"value":"Qwen 3.8 27B"},"story_url":{"matchLevel":"none","matchedWords":[],"value":"https://huggingface.co/Qwen/Qwen3.8-27B-FP8"}},"_tags":["comment","author_dofm","story_49299605"],"author":"dofm","children":[49309813],"comment_text":"Have you tried turning down the new Qwen&#x27;s reasoning effort level from xhigh, which it defaults at?<p>LM Studio isn&#x27;t exposing a dropdown for this, at least with the unsloth build.<p>Unsloth Studio &#x2F; Desktop does.","created_at":"2026-08-15T11:00:57Z","created_at_i":1786791657,"objectID":"49309556","parent_id":49306639,"story_id":49299605,"story_title":"Qwen 3.8 27B","story_url":"https://huggingface.co/Qwen/Qwen3.8-27B-FP8","updated_at":"2026-08-15T11:50:55Z"},{"_highlightResult":{"author":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["unsloth"],"value":"apex_<em>sloth</em>"},"comment_text":{"matchLevel":"none","matchedWords":[],"value":"Before I even look at a PR of sol I ask it to justify the loc. Quite often it comes back suggesting things it could simplify"},"story_title":{"matchLevel":"none","matchedWords":[],"value":"Why does Opus 5 feel worse to work with?"},"story_url":{"matchLevel":"none","matchedWords":[],"value":"https://mun-logadan.github.io/why-does-opus-5-feel-worse/"}},"_tags":["comment","author_apex_sloth","story_49296740"],"author":"apex_sloth","comment_text":"Before I even look at a PR of sol I ask it to justify the loc. Quite often it comes back suggesting things it could simplify","created_at":"2026-08-15T09:39:34Z","created_at_i":1786786774,"objectID":"49309138","parent_id":49306283,"story_id":49296740,"story_title":"Why does Opus 5 feel worse to work with?","story_url":"https://mun-logadan.github.io/why-does-opus-5-feel-worse/","updated_at":"2026-08-15T12:28:24Z"},{"_highlightResult":{"author":{"matchLevel":"none","matchedWords":[],"value":"dexterlagan"},"comment_text":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["unsloth"],"value":"It was in LMStudio (llama.cpp), Q4 by <em>Unsloth</em>. Applied the recommended defaults published by <em>Unsloth</em>."},"story_title":{"matchLevel":"none","matchedWords":[],"value":"Qwen 3.8 27B"},"story_url":{"matchLevel":"none","matchedWords":[],"value":"https://huggingface.co/Qwen/Qwen3.8-27B-FP8"}},"_tags":["comment","author_dexterlagan","story_49299605"],"author":"dexterlagan","comment_text":"It was in LMStudio (llama.cpp), Q4 by Unsloth. Applied the recommended defaults published by Unsloth.","created_at":"2026-08-15T08:09:00Z","created_at_i":1786781340,"objectID":"49308736","parent_id":49308293,"story_id":49299605,"story_title":"Qwen 3.8 27B","story_url":"https://huggingface.co/Qwen/Qwen3.8-27B-FP8","updated_at":"2026-08-16T10:44:42Z"},{"_highlightResult":{"author":{"matchLevel":"none","matchedWords":[],"value":"beltsazar"},"comment_text":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["unsloth"],"value":"Using a better quant Q6_K (<em>Unsloth</em>'s) compared to your quant Q4_K_M (LM Studio's) yields this: <a href=\"https://tools.simonwillison.net/markdown-svg-renderer#url=https%3A%2F%2Fgist.github.com%2Fdanieljl%2F0a942c0a6499d42be099117efe9a76cc\" rel=\"nofollow\">https://tools.simonwillison.net/markdown-svg-renderer#url=ht...</a><p>Chains exist. Red scarf is proactively added. (&quot;Maybe a scarf blowing in the wind for charm!) No hands/wings, though.<p>Generated 30.2k tokens in total and took 52 mins on M5 Pro in low power mode (it will possibly take less than half of that in auto energy mode)."},"story_title":{"matchLevel":"none","matchedWords":[],"value":"Qwen 3.8 27B"},"story_url":{"matchLevel":"none","matchedWords":[],"value":"https://huggingface.co/Qwen/Qwen3.8-27B-FP8"}},"_tags":["comment","author_beltsazar","story_49299605"],"author":"beltsazar","comment_text":"Using a better quant Q6_K (Unsloth&#x27;s) compared to your quant Q4_K_M (LM Studio&#x27;s) yields this: <a href=\"https:&#x2F;&#x2F;tools.simonwillison.net&#x2F;markdown-svg-renderer#url=https%3A%2F%2Fgist.github.com%2Fdanieljl%2F0a942c0a6499d42be099117efe9a76cc\" rel=\"nofollow\">https:&#x2F;&#x2F;tools.simonwillison.net&#x2F;markdown-svg-renderer#url=ht...</a><p>Chains exist. Red scarf is proactively added. (&quot;Maybe a scarf blowing in the wind for charm!) No hands&#x2F;wings, though.<p>Generated 30.2k tokens in total and took 52 mins on M5 Pro in low power mode (it will possibly take less than half of that in auto energy mode).","created_at":"2026-08-15T06:53:50Z","created_at_i":1786776830,"objectID":"49308368","parent_id":49304034,"story_id":49299605,"story_title":"Qwen 3.8 27B","story_url":"https://huggingface.co/Qwen/Qwen3.8-27B-FP8","updated_at":"2026-08-15T10:37:24Z"},{"_highlightResult":{"author":{"matchLevel":"none","matchedWords":[],"value":"MrDrMcCoy"},"comment_text":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["unsloth"],"value":"I gave you a direct link to where you can get a download, and explained why I chose to link you where I did. I can't post screenshots here drawing you a map of how to use a website. I don't work for Huggingface or any AI company and can affect no changes to how intuitive any of it is, and think is easy enough already.<p>Huggingface, <em>Unsloth</em>, and llama.cpp all have documentation you can follow that will exceed anything I can tell you here. LMstudio, Lemonade, or Ollama might be even easier for you to use. Take my suggestions or don't."},"story_title":{"matchLevel":"none","matchedWords":[],"value":"Qwen 3.8 27B"},"story_url":{"matchLevel":"none","matchedWords":[],"value":"https://huggingface.co/Qwen/Qwen3.8-27B-FP8"}},"_tags":["comment","author_MrDrMcCoy","story_49299605"],"author":"MrDrMcCoy","comment_text":"I gave you a direct link to where you can get a download, and explained why I chose to link you where I did. I can&#x27;t post screenshots here drawing you a map of how to use a website. I don&#x27;t work for Huggingface or any AI company and can affect no changes to how intuitive any of it is, and think is easy enough already.<p>Huggingface, Unsloth, and llama.cpp all have documentation you can follow that will exceed anything I can tell you here. LMstudio, Lemonade, or Ollama might be even easier for you to use. Take my suggestions or don&#x27;t.","created_at":"2026-08-15T06:17:56Z","created_at_i":1786774676,"objectID":"49308161","parent_id":49307713,"story_id":49299605,"story_title":"Qwen 3.8 27B","story_url":"https://huggingface.co/Qwen/Qwen3.8-27B-FP8","updated_at":"2026-08-15T08:00:09Z"},{"_highlightResult":{"author":{"matchLevel":"none","matchedWords":[],"value":"CMay"},"comment_text":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["unsloth"],"value":"Gemma 4 12B, Gemma 4 12B QAT, Gemma 4 31B, Gemma 4 31B QAT<p>Gemma 4 26BA4B would get close, but not quite and sometimes even get stuck in loops despite a repeat penalty.<p>Do not use any newer updated templates or <em>Unsloth</em> fixes.  Use older official templates that released with the models on the huggingface repo.  The template here worked: <a href=\"https://huggingface.co/google/gemma-4-12B-it/tree/657684fef0b5ac5d6bff39284ceb6ec3710b700e\" rel=\"nofollow\">https://huggingface.co/google/gemma-4-12B-it/tree/657684fef0...</a><p>llama-server --model &quot;model.gguf&quot; -fa on -np 1 --jinja --ctx-size 262144 -b 768 -ub 768 --cache-type-k f16 --cache-type-v q4_0 --repeat-penalty 1.1 --chat-template-file &quot;chat_template.jinja&quot;<p>If you don't explicitly point to the template file, then llama.cpp will either use the template inside the model file or it will use its own template copy and your results may vary.  Obviously some of the template fixes are useful to people, so it depends if you're having problems with tool calling or can't fix the tool calling in other ways for your scenario.<p>My experience with the QAT models was that quantizing v to q4_0 gave me better results than q8_0 or even f16.  I think the Gemma QAT models may have been QAT trained to expect a q4_0 quantized v cache.  If you're not using a QAT model, I would leave both at f16.<p>Another thing aside from using the QAT models and a Q4_0 v cache since you're having trouble fitting the models, is that you don't have to use the mmproj if you don't intend to use vision.  If you need vision, but are hurting on VRAM, then you should be using --no-mmproj-offload.  That will keep the mmproj loaded in system RAM instead of on your GPU.  Loading images will be a little bit slower, but it can still be quite fast and you'll have more breathing room on your GPU.  If you don't provide the mmproj file on the command line at all, then it won't load it anyway.  If you're using some program like LM Studio, a simple thing you can do is move the mmproj and mtp files out of the directory for the model so LM studio can't find them and then it won't load them at all.<p>For Qwen 3.8 27B, doing any quantizing definitely hurt results a lot, so in my case I used:\nllama-server --model &quot;Qwen3.8-27B-UD-Q4_K_XL.gguf&quot; --spec-type draft-mtp --spec-draft-p-min 0.35 --spec-draft-n-max 2 -fa on -np 1 --jinja --ctx-size 65536 -b 768 -ub 768 --cache-type-k f16 --cache-type-v f16"},"story_title":{"matchLevel":"none","matchedWords":[],"value":"Qwen 3.8 27B"},"story_url":{"matchLevel":"none","matchedWords":[],"value":"https://huggingface.co/Qwen/Qwen3.8-27B-FP8"}},"_tags":["comment","author_CMay","story_49299605"],"author":"CMay","comment_text":"Gemma 4 12B, Gemma 4 12B QAT, Gemma 4 31B, Gemma 4 31B QAT<p>Gemma 4 26BA4B would get close, but not quite and sometimes even get stuck in loops despite a repeat penalty.<p>Do not use any newer updated templates or Unsloth fixes.  Use older official templates that released with the models on the huggingface repo.  The template here worked: <a href=\"https:&#x2F;&#x2F;huggingface.co&#x2F;google&#x2F;gemma-4-12B-it&#x2F;tree&#x2F;657684fef0b5ac5d6bff39284ceb6ec3710b700e\" rel=\"nofollow\">https:&#x2F;&#x2F;huggingface.co&#x2F;google&#x2F;gemma-4-12B-it&#x2F;tree&#x2F;657684fef0...</a><p>llama-server --model &quot;model.gguf&quot; -fa on -np 1 --jinja --ctx-size 262144 -b 768 -ub 768 --cache-type-k f16 --cache-type-v q4_0 --repeat-penalty 1.1 --chat-template-file &quot;chat_template.jinja&quot;<p>If you don&#x27;t explicitly point to the template file, then llama.cpp will either use the template inside the model file or it will use its own template copy and your results may vary.  Obviously some of the template fixes are useful to people, so it depends if you&#x27;re having problems with tool calling or can&#x27;t fix the tool calling in other ways for your scenario.<p>My experience with the QAT models was that quantizing v to q4_0 gave me better results than q8_0 or even f16.  I think the Gemma QAT models may have been QAT trained to expect a q4_0 quantized v cache.  If you&#x27;re not using a QAT model, I would leave both at f16.<p>Another thing aside from using the QAT models and a Q4_0 v cache since you&#x27;re having trouble fitting the models, is that you don&#x27;t have to use the mmproj if you don&#x27;t intend to use vision.  If you need vision, but are hurting on VRAM, then you should be using --no-mmproj-offload.  That will keep the mmproj loaded in system RAM instead of on your GPU.  Loading images will be a little bit slower, but it can still be quite fast and you&#x27;ll have more breathing room on your GPU.  If you don&#x27;t provide the mmproj file on the command line at all, then it won&#x27;t load it anyway.  If you&#x27;re using some program like LM Studio, a simple thing you can do is move the mmproj and mtp files out of the directory for the model so LM studio can&#x27;t find them and then it won&#x27;t load them at all.<p>For Qwen 3.8 27B, doing any quantizing definitely hurt results a lot, so in my case I used:\nllama-server --model &quot;Qwen3.8-27B-UD-Q4_K_XL.gguf&quot; --spec-type draft-mtp --spec-draft-p-min 0.35 --spec-draft-n-max 2 -fa on -np 1 --jinja --ctx-size 65536 -b 768 -ub 768 --cache-type-k f16 --cache-type-v f16","created_at":"2026-08-15T06:16:33Z","created_at_i":1786774593,"objectID":"49308155","parent_id":49307727,"story_id":49299605,"story_title":"Qwen 3.8 27B","story_url":"https://huggingface.co/Qwen/Qwen3.8-27B-FP8","updated_at":"2026-08-16T10:40:27Z"},{"_highlightResult":{"author":{"matchLevel":"none","matchedWords":[],"value":"zenoprax"},"comment_text":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["unsloth"],"value":"Depends on your tooling and quant? I grabbed the <em>unsloth</em> Q3 and it works out if the box in opencode. I had issues with OpenWebUI with a random 3.6 A3B."},"story_title":{"matchLevel":"none","matchedWords":[],"value":"Qwen 3.8 27B"},"story_url":{"matchLevel":"none","matchedWords":[],"value":"https://huggingface.co/Qwen/Qwen3.8-27B-FP8"}},"_tags":["comment","author_zenoprax","story_49299605"],"author":"zenoprax","comment_text":"Depends on your tooling and quant? I grabbed the unsloth Q3 and it works out if the box in opencode. I had issues with OpenWebUI with a random 3.6 A3B.","created_at":"2026-08-15T02:36:53Z","created_at_i":1786761413,"objectID":"49307069","parent_id":49301556,"story_id":49299605,"story_title":"Qwen 3.8 27B","story_url":"https://huggingface.co/Qwen/Qwen3.8-27B-FP8","updated_at":"2026-08-15T02:38:08Z"},{"_highlightResult":{"author":{"matchLevel":"none","matchedWords":[],"value":"MrDrMcCoy"},"comment_text":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["unsloth"],"value":"&gt; &quot;I went to youtube, gmail, ycombinator, deliveroo, then I went to another site I randomly chose, because they're the best, duh&quot;<p>The path I described is all within HuggingFace.<p>&gt; And How am I supposed to know that arriving to huggingface as a new user? Enlighten me.<p>Because I told you, knowing it was the best starting point for newbies.<p>&gt; Cool.. Why don't they share em because I genuinely cant find em, I'm dumb.<p>You could go to Qwen's organization page on HuggingFace, it has a search function at the top, but you would be better served sticking with <em>Unsloth</em>.<p>&gt; Well, I'm willing, but not from people who I might burn good will. Gracious. You do you.<p>Expecting others to do everything for you is not the same as trying things and asking questions about what you found."},"story_title":{"matchLevel":"none","matchedWords":[],"value":"Qwen 3.8 27B"},"story_url":{"matchLevel":"none","matchedWords":[],"value":"https://huggingface.co/Qwen/Qwen3.8-27B-FP8"}},"_tags":["comment","author_MrDrMcCoy","story_49299605"],"author":"MrDrMcCoy","children":[49307713,49307669],"comment_text":"&gt; &quot;I went to youtube, gmail, ycombinator, deliveroo, then I went to another site I randomly chose, because they&#x27;re the best, duh&quot;<p>The path I described is all within HuggingFace.<p>&gt; And How am I supposed to know that arriving to huggingface as a new user? Enlighten me.<p>Because I told you, knowing it was the best starting point for newbies.<p>&gt; Cool.. Why don&#x27;t they share em because I genuinely cant find em, I&#x27;m dumb.<p>You could go to Qwen&#x27;s organization page on HuggingFace, it has a search function at the top, but you would be better served sticking with Unsloth.<p>&gt; Well, I&#x27;m willing, but not from people who I might burn good will. Gracious. You do you.<p>Expecting others to do everything for you is not the same as trying things and asking questions about what you found.","created_at":"2026-08-15T01:43:11Z","created_at_i":1786758191,"objectID":"49306803","parent_id":49306691,"story_id":49299605,"story_title":"Qwen 3.8 27B","story_url":"https://huggingface.co/Qwen/Qwen3.8-27B-FP8","updated_at":"2026-08-15T04:46:53Z"},{"_highlightResult":{"author":{"matchLevel":"none","matchedWords":[],"value":"rasengan0"},"comment_text":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["unsloth"],"value":"Got Qwen3.8 to run on my Framework 12 Intel Core 13 Gen Raptor Lake i5-1334U small laptop with 48G RAM stick:<p>llama serve -hf <em>unsloth</em>/Qwen3.8-27B-GGUF:UD-Q4_K_XL<p>but it failed my basic prompt to compose a vim regex to match CamelCaseWords<p>downgraded a bit with Q4_K_M from ollama run qwen3.8:27b<p>and /set nothink and at least 1 regex matched FooBar<p>prompt eval at 2.8 t/s\neval at 0.94t/s<p>I particularly enjoyed this usage of the regex: \n/%\\1\\%/ ... onward for 700+ characters of \\%\\/\n:-)"},"story_title":{"matchLevel":"none","matchedWords":[],"value":"Qwen 3.8 27B"},"story_url":{"matchLevel":"none","matchedWords":[],"value":"https://huggingface.co/Qwen/Qwen3.8-27B-FP8"}},"_tags":["comment","author_rasengan0","story_49299605"],"author":"rasengan0","comment_text":"Got Qwen3.8 to run on my Framework 12 Intel Core 13 Gen Raptor Lake i5-1334U small laptop with 48G RAM stick:<p>llama serve -hf unsloth&#x2F;Qwen3.8-27B-GGUF:UD-Q4_K_XL<p>but it failed my basic prompt to compose a vim regex to match CamelCaseWords<p>downgraded a bit with Q4_K_M from ollama run qwen3.8:27b<p>and &#x2F;set nothink and at least 1 regex matched FooBar<p>prompt eval at 2.8 t&#x2F;s\neval at 0.94t&#x2F;s<p>I particularly enjoyed this usage of the regex: \n&#x2F;%\\1\\%&#x2F; ... onward for 700+ characters of \\%\\&#x2F;\n:-)","created_at":"2026-08-15T01:39:33Z","created_at_i":1786757973,"objectID":"49306779","parent_id":49299963,"story_id":49299605,"story_title":"Qwen 3.8 27B","story_url":"https://huggingface.co/Qwen/Qwen3.8-27B-FP8","updated_at":"2026-08-15T10:54:24Z"},{"_highlightResult":{"author":{"matchLevel":"none","matchedWords":[],"value":"bilekas"},"comment_text":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["unsloth"],"value":"&gt; I went to Huggingface, went to the <em>Unsloth</em> org, as they tend to be the best, went to the Model page, and went to the &quot;Files and versions&quot; tab.<p>&quot;I went to youtube, gmail, ycombinator, deliveroo, then I went to another site I randomly chose, because they're the best, duh&quot;<p>Okay.<p>&gt; <em>Unsloth</em> AI is a very popular, highly reputable organization that takes upstream model files<p>And How am I supposed to know that arriving to huggingface as a new user? Enlighten me.<p>&gt; Qwen also provides models in GGUF format on Huggingface<p>Cool.. Why don't they share em because I genuinely cant find em, I'm dumb.<p>&gt; Your attitude and unwillingness to even try and learn on your own when people have tried helping have burned my good will<p>Well, I'm willing, but not from people who I might burn good will. Gracious. You do you."},"story_title":{"matchLevel":"none","matchedWords":[],"value":"Qwen 3.8 27B"},"story_url":{"matchLevel":"none","matchedWords":[],"value":"https://huggingface.co/Qwen/Qwen3.8-27B-FP8"}},"_tags":["comment","author_bilekas","story_49299605"],"author":"bilekas","children":[49306803],"comment_text":"&gt; I went to Huggingface, went to the Unsloth org, as they tend to be the best, went to the Model page, and went to the &quot;Files and versions&quot; tab.<p>&quot;I went to youtube, gmail, ycombinator, deliveroo, then I went to another site I randomly chose, because they&#x27;re the best, duh&quot;<p>Okay.<p>&gt; Unsloth AI is a very popular, highly reputable organization that takes upstream model files<p>And How am I supposed to know that arriving to huggingface as a new user? Enlighten me.<p>&gt; Qwen also provides models in GGUF format on Huggingface<p>Cool.. Why don&#x27;t they share em because I genuinely cant find em, I&#x27;m dumb.<p>&gt; Your attitude and unwillingness to even try and learn on your own when people have tried helping have burned my good will<p>Well, I&#x27;m willing, but not from people who I might burn good will. Gracious. You do you.","created_at":"2026-08-15T01:28:21Z","created_at_i":1786757301,"objectID":"49306691","parent_id":49306541,"story_id":49299605,"story_title":"Qwen 3.8 27B","story_url":"https://huggingface.co/Qwen/Qwen3.8-27B-FP8","updated_at":"2026-08-15T01:43:53Z"},{"_highlightResult":{"author":{"matchLevel":"none","matchedWords":[],"value":"MrDrMcCoy"},"comment_text":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["unsloth"],"value":"&gt; Excuse me, but thats a direct link you've just sent. I asked where I can find the links. I like to believe in the source of truth.<p>You asked where to find the GGUF files of this model for direct download and I provided it. Almost all useful model files that can be downloaded are hosted on Huggingface.<p>&gt; They shared a lot of links, I'm struggling to find yours. Where did yours come from ?<p>I went to Huggingface, went to the <em>Unsloth</em> org, as they tend to be the best, went to the Model page, and went to the &quot;Files and versions&quot; tab.<p>&gt; Who is <em>Unsloth</em> AI? Have they modified the model ? Is this really the source of truth ?<p><em>Unsloth</em> AI is a very popular, highly reputable organization that takes upstream model files, performs some optimization, and provides models in various formats. Apart from speed tweaks, they do not modify the models. They also provide useful benchmarks, copious documentation for local execution, and a Studio application for easy execution and post-training of models.<p>&gt; Do you see how steep the barrier for entry is to do anything right ?<p>No. Searching for this information is not difficult. The llama.cpp documentation and guides that <em>Unsloth</em> provide are all you need. Search engines can take you further if you want.<p>&gt; <em>unsloth</em>ai is not a name qwen has ever used. So you're sharing a link to a model that isn't from the owner, while saying it's the owner's. I'm not comfortable with that<p>Qwen also provides models in GGUF format on Huggingface, but they will not be as performant. Even when first-party GGUFs are available, most people will prefer quants from <em>Unsloth</em> or a few other popular optimizer accounts.<p>&gt; I want AI to be a better tool.<p>Best of luck. Your attitude and unwillingness to even try and learn on your own when people have tried helping have burned my good will, and this is as far as I'm willing to carry you."},"story_title":{"matchLevel":"none","matchedWords":[],"value":"Qwen 3.8 27B"},"story_url":{"matchLevel":"none","matchedWords":[],"value":"https://huggingface.co/Qwen/Qwen3.8-27B-FP8"}},"_tags":["comment","author_MrDrMcCoy","story_49299605"],"author":"MrDrMcCoy","children":[49306691],"comment_text":"&gt; Excuse me, but thats a direct link you&#x27;ve just sent. I asked where I can find the links. I like to believe in the source of truth.<p>You asked where to find the GGUF files of this model for direct download and I provided it. Almost all useful model files that can be downloaded are hosted on Huggingface.<p>&gt; They shared a lot of links, I&#x27;m struggling to find yours. Where did yours come from ?<p>I went to Huggingface, went to the Unsloth org, as they tend to be the best, went to the Model page, and went to the &quot;Files and versions&quot; tab.<p>&gt; Who is Unsloth AI? Have they modified the model ? Is this really the source of truth ?<p>Unsloth AI is a very popular, highly reputable organization that takes upstream model files, performs some optimization, and provides models in various formats. Apart from speed tweaks, they do not modify the models. They also provide useful benchmarks, copious documentation for local execution, and a Studio application for easy execution and post-training of models.<p>&gt; Do you see how steep the barrier for entry is to do anything right ?<p>No. Searching for this information is not difficult. The llama.cpp documentation and guides that Unsloth provide are all you need. Search engines can take you further if you want.<p>&gt; unslothai is not a name qwen has ever used. So you&#x27;re sharing a link to a model that isn&#x27;t from the owner, while saying it&#x27;s the owner&#x27;s. I&#x27;m not comfortable with that<p>Qwen also provides models in GGUF format on Huggingface, but they will not be as performant. Even when first-party GGUFs are available, most people will prefer quants from Unsloth or a few other popular optimizer accounts.<p>&gt; I want AI to be a better tool.<p>Best of luck. Your attitude and unwillingness to even try and learn on your own when people have tried helping have burned my good will, and this is as far as I&#x27;m willing to carry you.","created_at":"2026-08-15T01:09:51Z","created_at_i":1786756191,"objectID":"49306541","parent_id":49305591,"story_id":49299605,"story_title":"Qwen 3.8 27B","story_url":"https://huggingface.co/Qwen/Qwen3.8-27B-FP8","updated_at":"2026-08-15T07:59:38Z"},{"_highlightResult":{"author":{"matchLevel":"none","matchedWords":[],"value":"jakswa"},"comment_text":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["unsloth"],"value":"I'm in the exact same boat with a 7900 XT and a good Glimmer 30B experience. I was really hoping qwen 3.8 would bring some memory/space efficiency savings along the lines of whatever is going on with Glimmer 30B. I have been surprised that a 30 billion model fits and runs better (at higher <em>unsloth</em> quantization! UD-Q4_K_XL fits!) than a 27 billion model."},"story_title":{"matchLevel":"none","matchedWords":[],"value":"Qwen 3.8 27B"},"story_url":{"matchLevel":"none","matchedWords":[],"value":"https://huggingface.co/Qwen/Qwen3.8-27B-FP8"}},"_tags":["comment","author_jakswa","story_49299605"],"author":"jakswa","children":[49308794],"comment_text":"I&#x27;m in the exact same boat with a 7900 XT and a good Glimmer 30B experience. I was really hoping qwen 3.8 would bring some memory&#x2F;space efficiency savings along the lines of whatever is going on with Glimmer 30B. I have been surprised that a 30 billion model fits and runs better (at higher unsloth quantization! UD-Q4_K_XL fits!) than a 27 billion model.","created_at":"2026-08-15T00:20:53Z","created_at_i":1786753253,"objectID":"49306222","parent_id":49304616,"story_id":49299605,"story_title":"Qwen 3.8 27B","story_url":"https://huggingface.co/Qwen/Qwen3.8-27B-FP8","updated_at":"2026-08-15T08:19:09Z"},{"_highlightResult":{"author":{"matchLevel":"none","matchedWords":[],"value":"suprjami"},"comment_text":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["unsloth"],"value":"You have understood correctly.<p>One really would think these companies (including Google) who spend many millions of dollars on compute could write a few hundred lines of Jinja correctly, so their investment works optimally or at all.<p>But they don't.<p>Then a couple of individuals on HuggingFace fix it, either a 2-person startup like <em>Unsloth</em> or a volunteer like froggeric.<p>I also don't understand how this repeatedly happens."},"story_title":{"matchLevel":"none","matchedWords":[],"value":"Qwen 3.8 27B"},"story_url":{"matchLevel":"none","matchedWords":[],"value":"https://huggingface.co/Qwen/Qwen3.8-27B-FP8"}},"_tags":["comment","author_suprjami","story_49299605"],"author":"suprjami","children":[49307739],"comment_text":"You have understood correctly.<p>One really would think these companies (including Google) who spend many millions of dollars on compute could write a few hundred lines of Jinja correctly, so their investment works optimally or at all.<p>But they don&#x27;t.<p>Then a couple of individuals on HuggingFace fix it, either a 2-person startup like Unsloth or a volunteer like froggeric.<p>I also don&#x27;t understand how this repeatedly happens.","created_at":"2026-08-14T23:17:19Z","created_at_i":1786749439,"objectID":"49305761","parent_id":49301680,"story_id":49299605,"story_title":"Qwen 3.8 27B","story_url":"https://huggingface.co/Qwen/Qwen3.8-27B-FP8","updated_at":"2026-08-16T05:30:11Z"},{"_highlightResult":{"author":{"matchLevel":"none","matchedWords":[],"value":"bilekas"},"comment_text":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["unsloth"],"value":"Excuse me, but thats a direct link you've just sent. I asked where I can find the links. I like to believe in the source of truth.<p>They shared a lot of links, I'm struggling to find yours. Where did yours come from ?<p>Who is <em>Unsloth</em> AI? Have they modified the model ? Is this really the source of truth ?<p>Do you see how steep the barrier for entry is to do anything right ?<p><em>unsloth</em>ai is not a name qwen has ever used. So you're sharing a link to a model that isn't from the owner, while saying it's the owner's. I'm not comfortable with that, and I want AI to be a better tool."},"story_title":{"matchLevel":"none","matchedWords":[],"value":"Qwen 3.8 27B"},"story_url":{"matchLevel":"none","matchedWords":[],"value":"https://huggingface.co/Qwen/Qwen3.8-27B-FP8"}},"_tags":["comment","author_bilekas","story_49299605"],"author":"bilekas","children":[49306541],"comment_text":"Excuse me, but thats a direct link you&#x27;ve just sent. I asked where I can find the links. I like to believe in the source of truth.<p>They shared a lot of links, I&#x27;m struggling to find yours. Where did yours come from ?<p>Who is Unsloth AI? Have they modified the model ? Is this really the source of truth ?<p>Do you see how steep the barrier for entry is to do anything right ?<p>unslothai is not a name qwen has ever used. So you&#x27;re sharing a link to a model that isn&#x27;t from the owner, while saying it&#x27;s the owner&#x27;s. I&#x27;m not comfortable with that, and I want AI to be a better tool.","created_at":"2026-08-14T22:52:26Z","created_at_i":1786747946,"objectID":"49305591","parent_id":49305564,"story_id":49299605,"story_title":"Qwen 3.8 27B","story_url":"https://huggingface.co/Qwen/Qwen3.8-27B-FP8","updated_at":"2026-08-15T01:10:23Z"},{"_highlightResult":{"author":{"matchLevel":"none","matchedWords":[],"value":"MrDrMcCoy"},"comment_text":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["unsloth"],"value":"Here: <a href=\"https://huggingface.co/unsloth/Qwen3.8-27B-GGUF/tree/main\" rel=\"nofollow\">https://huggingface.co/<em>unsloth</em>/Qwen3.8-27B-GGUF/tree/main</a>"},"story_title":{"matchLevel":"none","matchedWords":[],"value":"Qwen 3.8 27B"},"story_url":{"matchLevel":"none","matchedWords":[],"value":"https://huggingface.co/Qwen/Qwen3.8-27B-FP8"}},"_tags":["comment","author_MrDrMcCoy","story_49299605"],"author":"MrDrMcCoy","children":[49305591],"comment_text":"Here: <a href=\"https:&#x2F;&#x2F;huggingface.co&#x2F;unsloth&#x2F;Qwen3.8-27B-GGUF&#x2F;tree&#x2F;main\" rel=\"nofollow\">https:&#x2F;&#x2F;huggingface.co&#x2F;unsloth&#x2F;Qwen3.8-27B-GGUF&#x2F;tree&#x2F;main</a>","created_at":"2026-08-14T22:48:49Z","created_at_i":1786747729,"objectID":"49305564","parent_id":49304891,"story_id":49299605,"story_title":"Qwen 3.8 27B","story_url":"https://huggingface.co/Qwen/Qwen3.8-27B-FP8","updated_at":"2026-08-14T22:53:22Z"}],"hitsPerPage":20,"nbHits":7174,"nbPages":50,"page":0,"params":"query=unsloth&advancedSyntax=true&analyticsTags=backend","processingTimeMS":8,"processingTimingsMS":{"_request":{"roundTrip":16},"afterFetch":{"format":{"total":1}},"fetch":{"query":6,"total":7},"total":8},"query":"unsloth","serverTimeMS":10}
