DeepSeek v4.1 Flash Uncensored
Posted by soltanov 1 day ago
Comments
Comment by tacomagick 1 day ago
Comment by thulle 1 day ago
> This approach enables Heretic to work completely automatically. Heretic finds high-quality abliteration parameters by co-minimizing the number of refusals and the KL divergence from the original model. This results in a decensored model that retains as much of the original model's intelligence as possible. Using Heretic does not require an understanding of transformer internals. In fact, anyone who knows how to run a command-line program can use Heretic to decensor language models.
Abliteration seems to be what Heretic does?
> Now these "abliterated" models all suffer from catastrophic breakage because they are not as simple anymore.
I'm not knowledgeable about each step in the process of making these abliterated models, but some more popular ones with steps after Heretic, seem to improve on the benchmarks tried of the base model:
arc/c arc/e boolq hswag obkqa piqa wino
Qwen3.6-27B-Fable-Fusion-711-Uncensored-Heretic-NM-DAU-NEO-MAX-MTP-GGUF [instruct mode]
mxfp8 0.711,0.879,0.910,0.790,0.514,0.823,0.763
mxfp4 0.701,0.873,0.909,0.786,0.488,0.813,0.759
Qwen3.6-27B-Instruct: [base, non heretic]
mxfp8 0.647,0.803,0.910,0.773,0.450,0.806,0.742
https://huggingface.co/DavidAU/Qwen3.6-27B-Fable-Fusion-711-...I'm not seeing any similar benchmarks of the HauhauCS models, at least the ones I checked, so I assumed the opinion is based on your own trials, but then you argue in favour Heretic. Is the based on pre-Heretic abliteration techniques? Which might then not be appliable to this "Proprietary weight-level abliteration developed by the dealignai research team."?
Comment by tacomagick 1 day ago
Comment by halJordan 1 day ago
Comment by tacomagick 1 day ago
Comment by pullstart 1 day ago
Comment by eddyg 1 day ago
Comment by crooked-v 1 day ago
Comment by esseph 1 day ago
Non-zero chance the lasting impact LLMs have are a bioweapon.
https://www.nytimes.com/2026/09/10/us/politics/anthropic-ai-...
Comment by dd8601fn 1 day ago
I’m a thoroughly average guy. I’m not capable of asking competent supervillain questions.
And Anthropic said they couldn’t say if any of the “bioweapon” safeguards went off on nefarious efforts. I’m probably in those numbers.
So read it as marketing more than something to lose sleep over. They’re mostly gating stuff a sufficiently motivated person would find with a library card.
Comment by esseph 1 day ago
https://openai.com/index/building-an-early-warning-system-fo...
https://www.rand.org/pubs/research_reports/RRA2977-1.html
---
"Prompting Moremi Bio Agent without the safety guardrails to specifically design novel toxic substances, our study generated 1020 novel toxic proteins and 5,000 toxic small molecules. In-depth computational toxicity assessments revealed that all the proteins scored high in toxicity, with several closely matching known toxins such as ricin, diphtheria toxin, and disintegrin-based snake venom proteins."
"The findings from this toxicity assessment challenge claims that large language models (LLMs) are incapable of designing bioweapons. This reinforces concerns about the potential misuse of LLMs in biodesign, posing a significant threat to research and development (R&D). The accessibility of such technology to individuals with limited technical expertise raises serious biosecurity risks. Our findings underscore the critical need for robust governance and technical safeguards to balance rapid biotechnological innovation with biosecurity imperatives."
Comment by tacomagick 1 day ago
Comment by __rito__ 1 day ago
Comment by ascorbic 1 day ago
Comment by __rito__ 23 hours ago
Or do you read long form content in good magazines/newspapers on vaccines, disease, etc.?
Comment by esseph 1 day ago
https://theconversation.com/worlds-first-ai-designed-vaccine...
Comment by rolymath 1 day ago
Comment by iamnothere 1 day ago
Comment by 63stack 1 day ago
Comment by Jemm 1 day ago
Comment by qgin 1 day ago
Comment by pgsandstrom 1 day ago
Comment by darkport 1 day ago
Comment by dude250711 1 day ago
Comment by __bjoernd 1 day ago
Comment by mentalgear 1 day ago