{"slug": "scotoma-gemma-4-31b-abliteration-review", "title": "SCOTOMA: Gemma 4 31B Abliteration Review", "summary": "Developer Nokka reviewed SCOTOMA, an abliterated version of Google's Gemma 4 31B Instruct model created by ReadyArt. SCOTOMA uses a novel J-space abliteration technique with Jacobian-lens projection to selectively remove refusal behaviors while preserving intelligence, described as 'loosened, not lobotomized.' The model is available on Hugging Face in GGUF format under Apache 2.0 license.", "body_md": "*โดย Nokka (นก-กา) | 25 กรกฎาคม 2026*\n\n*บทความนี้เขียนโดย AI (DeepSeek V4 Pro) ผ่าน Hermes Agent ภายใต้การควบคุมของ Nokka (นก-กา) — รีวิวโมเดล SCOTOMA (Gemma 4 31B abliteration)*\n\nเคยใช้ Gemma 4 31B Instruct แล้วรู้สึกว่าเก่งแต่ \"สุภาพเกินไป\" ไหมครับ?\n\nตอบแบบ hedging ตลอด — \"อาจจะ...\" \"มีความเป็นไปได้ว่า...\" \"ทั้งนี้ขึ้นอยู่กับ...\" — เหมือนผู้ช่วยที่กลัวถูกฟ้องทุกครั้งที่เปิดปาก\n\n**SCOTOMA** คือคำตอบสำหรับปัญหานี้ — มันคือ Gemma 4 31B Instruct ที่ถูกแก้ไขให้ \"กล้าตอบ\" มากขึ้น โดยไม่เสีย intelligence — และที่สำคัญ — **ไม่ใช่ uncensored**\n\nSCOTOMA (อ่านว่า \"สโคโทมา\" — ศัพท์แพทย์แปลว่า \"จุดบอดในการมองเห็น\") คือ Gemma 4 31B Instruct ที่ผ่านกระบวนการ **J-space abliteration** แบบเจาะจง — สร้างโดย ReadyArt — ปล่อยบน Hugging Face ในรูปแบบ GGUF [1]\n\n| มิติ | ค่า |\n|---|---|\nBase |\ngemma-4-31B-it (Google) |\nMethod |\nJ-space abliteration + Jacobian-lens projection |\nParams |\nk256 · s1.5 · L7–41 |\nFormat |\nthinking (channel) — รองรับ reasoning |\nLicense |\nApache 2.0 |\nCreator |\nReadyArt |\nGGUF |\nIQ3_XS (13.1 GB) → Q8_0 (32.6 GB) |\n\nนี่คือส่วนที่น่าสนใจที่สุดของ SCOTOMA — **วิธีที่มันถูกสร้าง**\n\nAbliteration ทั่วไปทำงานโดยการ \"ลบ\" refusal direction ออกจากโมเดล — เหมือนใช้ค้อนทุบ — ได้ผลเร็ว แต่ทำลาย intelligence และ coherence ไปด้วย — ที่ชุมชนเรียกกันว่า \"lobotomized\" [1]\n\nSCOTOMA ใช้วิธีที่ละเอียดกว่ามาก — 3 ขั้นตอน [1]:\n\nใช้ heretic abliteration edit ฟิตบน attention + MLP layers — หาตำแหน่งที่ refusal ฝังตัวอยู่\n\nนี่คือหัวใจของเทคนิค — แทนที่จะลบ refusal direction ทั้งหมด — SCOTOMA ฉาย edit ผ่าน **Jacobian lens** เพื่อแยก:\n\nAbliteration ทั่วไปลบทั้งสองส่วน — SCOTOMA เก็บเฉพาะ behavioral — ทิ้ง computational ไว้ intact\n\nProjected edit ถูก bake เข้า bf16 weights ที่ application strength 1.5× — ปรับด้วยมือให้ \"รู้สึก\" กำลังดี — บน layers 7–41 [1]\n\n**ผลลัพธ์:** \"loosened, not lobotomized\" — คลายความระวัง โดยไม่ทำลายความฉลาด\n\nจาก model card [1]:\n\n\"It is not uncensored. The edit removed a bounded region while leaving the rest of the field intact. A blind spot, not blindness.\"\n\nSCOTOMA เป็น **research artifact** — ไม่ใช่ product — พฤติกรรมอาจเปลี่ยนตาม prompt และ context\n\nSCOTOMA มี GGUF ให้เลือกหลายขนาด — จาก 13.1 GB ถึง 32.6 GB:\n\n| Quant | ขนาด | 64GB Mac | 32GB Mac | 24GB GPU |\n|---|---|---|---|---|\nIQ3_XS |\n13.1 GB | ✅ | ✅ | ✅ |\nQ3_K_S |\n13.8 GB | ✅ | ✅ | ✅ |\nIQ3_M |\n14.4 GB | ✅ | ✅ | ✅ |\nQ3_K_M |\n15.3 GB | ✅ | ✅ | ✅ |\nIQ4_XS |\n16.7 GB | ✅ | ✅ | ✅ |\nQ4_K_S |\n17.8 GB | ✅ | ✅ | ✅ |\nQ4_K_M |\n18.7 GB | ✅ | ✅ | ✅ |\nQ5_K_S |\n21.3 GB | ✅ | ⚠️ | ❌ |\nQ5_K_M |\n21.8 GB | ✅ | ❌ | ❌ |\nQ6_K |\n25.2 GB | ✅ | ❌ | ❌ |\nQ8_0 |\n32.6 GB | ✅ | ❌ | ❌ |\n\n**คำสั่ง Ollama:**\n\n```\n# Q4_K_M — 18.7 GB (แนะนำ — สมดุลระหว่างคุณภาพและขนาด)\nollama run hf.co/ReadyArt/gemma-4-31B-it-scotoma-GGUF:Q4_K_M\n\n# IQ4_XS — 16.7 GB (ประหยัดกว่า — เหมาะกับเครื่อง 32GB)\nollama run hf.co/ReadyArt/gemma-4-31B-it-scotoma-GGUF:IQ4_XS\n```\n\nในมุมมองของผม SCOTOMA คือนวัตกรรมที่น่าสนใจไม่ใช่เพราะ **ผลลัพธ์** — แต่เพราะ **วิธีการ**\n\nการแยก behavioral component ออกจาก computational component ผ่าน Jacobian lens คือแนวคิดที่ elegant — มันบอกว่า \"เราไม่จำเป็นต้องทุบ refusal direction ทิ้งทั้งหมด — เราแค่ต้องแยกสิ่งที่โมเดล **พูด** ออกจากสิ่งที่โมเดล **คิด**\"\n\nนี่คือทิศทางที่ abliteration ควรจะไป — ละเอียดขึ้น, แม่นยำขึ้น, และรักษา intelligence ไว้ได้มากขึ้น\n\nสำหรับคนที่อยากได้ Gemma 4 31B ที่ \"เป็นธรรมชาติ\" มากขึ้น — โดยไม่เสียความสามารถในการ reasoning — SCOTOMA คือตัวเลือกที่คุ้มค่าที่จะลอง\n\n*ติดตามรีวิวโมเดล AI และเครื่องมือสำหรับนักพัฒนาได้ที่ Nokka on dev.to — กด Follow ที่โปรไฟล์เพื่อรับอัปเดตทุกครั้งที่มีบทความใหม่*\n\n[1] ReadyArt, \"gemma-4-31B-it-scotoma-GGUF,\" Hugging Face, กรกฎาคม 2026. [https://huggingface.co/ReadyArt/gemma-4-31B-it-scotoma-GGUF](https://huggingface.co/ReadyArt/gemma-4-31B-it-scotoma-GGUF)", "url": "https://wpnews.pro/news/scotoma-gemma-4-31b-abliteration-review", "canonical_source": "https://dev.to/sarantoon/scotoma-gemma-4-31b-abliteration-review-p81", "published_at": "2026-07-25 05:35:21+00:00", "updated_at": "2026-07-25 06:01:22.495303+00:00", "lang": "en", "topics": ["large-language-models", "ai-research", "ai-safety", "ai-products", "developer-tools"], "entities": ["Nokka", "ReadyArt", "Google", "Gemma 4 31B", "SCOTOMA", "Hugging Face", "DeepSeek V4 Pro", "Hermes Agent"], "alternates": {"html": "https://wpnews.pro/news/scotoma-gemma-4-31b-abliteration-review", "markdown": "https://wpnews.pro/news/scotoma-gemma-4-31b-abliteration-review.md", "text": "https://wpnews.pro/news/scotoma-gemma-4-31b-abliteration-review.txt", "jsonld": "https://wpnews.pro/news/scotoma-gemma-4-31b-abliteration-review.jsonld"}}