• 0 posts
  • 1 comment
Joined 1 year ago
Cake day: September 12th, 2025
  • I’ll try to answer as best as I can, but I can be barely be called a hobbyist, so hopefully someone else can give a better answer.

    1: Unfortunately, pretty much all models have a liberal slant to them, including the Chinese ones. Partly because of the nature of the training data (especially in English), and because no major organisation has actually attempted to create a socialist model, as far as I’m aware.

    As for your hardware, I think Qwen3.8-27B (from Alibaba) would fit comfortably at reasonable quantisation, though it might be a bit slow. The last gen Qwen3.6-35B-A3B would likely be significantly faster, though also less capable. If you don’t mind US models, Google’s Gemma-4 series is also pretty good, with the 31B and 26B-A4 models likely most suited for your hardware

    2 & 3: The next step up would be Qwen3.8-Flash-Next, which would probably want at least 90GB of RAM to run decently, further up you have GLM 3.5 Flash, which you would want >200 GB of RAM for. I don’t know how much VRAM you need to run these models at a decent speed though. In general, getting more RAM would let you run bigger models, while beefier GPUs with more VRAM would let you run them faster.

    4: Not quite sure what your needs are in terms of tuning. These days, a good system prompt is usually enough as models are pretty good generalists now. Sometimes they will self-censor, so you might want an abliterated version, which is the same model with refusals stripped out. If you really have to do finetuning, I think Unsloth has a tool for training LoRAs or QLoRAs for a model with consumer hardware, though my information is probably quite outdated as I haven’t paid attention to local fine-tuning in a while.

    Anyway, please don’t take my words as gospel, I can barely run anything myself anyway given my 16GB GPU, so I’m definitely not qualified to offer the best answer. I also don’t consider myself much of a techie anyway. Hopefully this helps though.