Cookbook Model List #1962
Replies: 15 comments 4 replies
|
i think one of the fine tuned 'uncencored" ai should be added: |
|
I think lfm2.5 models should be added its better then most qwen models around its weights also running on m1 air base model 8gb ram 7gpu 8cpu 256 gb the very budget friendly MacBook Air m1 sold from Walmart |
|
|
Qwopus models are really nice, qwen trained on opus traces; |
|
Good uncensored Gemma 4 model that won't stop working cause it encounters spicy words >_> |
|
Gemma 4 12B
Blog post: https://blog.google/innovation-and-ai/technology/developers-tools/introducing-gemma-4-12B/ |
|
Gemma 4 12B
|
|
Qwen3.5 Mixture of Quants + Multi-Token-Prediction
Specifically moq-3.6, runs on 8GB Vram, fast generation, modern model with quantization that is 97% accuracy against baseline. A generally good model for constrained hardware. |
|
HI, |
|
https://huggingface.co/deucebucket/Qwen3.6-35B-A3B-Cerebellum-GGUF https://huggingface.co/deucebucket/Gemma-4-26B-A4B-it-Heretic-Cerebellum-GGUF https://huggingface.co/deucebucket/Qwen3.6-27B-Cerebellum-GGUF |
|
put uncensored to model filter |
|
Hf Link: https://huggingface.co/zai-org/GLM-5.2 |
|
|
Worth looking at both the Agents A1 project as well as the Qwen Agentworld development. Both came out as recent pushes from official qwen team as well as a science lab for what's best for agents (over the last 2 weeks) Submitted pewdiepie-archdaemon/odysseus/discussions/5493 on this, as well as it being a fully curated vllm command guide for WSL 5090 owners. I opt for debian myself, as it's fairly straightforward. And, minimalist windows if you're really privacy-focused these days puts in some decent work on the privacy front, especially if you don't mind group policy and correct opts in telemetry. So a mini cookbook is on that discussion page for any card with over 24gb vram. Might have to adjust the quant slightly (int3 or something) but int4 is what I use. Would need two 5090s for fp8. Agents A1 has documentation about the science lab here 9:10 in the video |
|
@pewdiepie-archdaemon models I have configured with Odysseus and tested as working with tool use: https://huggingface.co/zai-org/GLM-5.2 BF16 non quantized, and quantized FP8 and Int4/Int8 variants.
Should have Mistral Large 3 675B FP8 tires kicked in the next week or so. I have been exploring agentic ops/sre assist using local models similar to what Anthropic and OpenAI do with Claude.ai and chatgpt. |
Uh oh!
There was an error while loading. Please reload this page.
Uh oh!
There was an error while loading. Please reload this page.
Cookbook is a curated list of models.
If you have a request for models missing, please leave a comment below:
Thank you
All reactions