
มีคนเอา Qwen3.5b มา fine tune เป็น Ornith ซึ่งหลายๆคนชมว่าฉลาดขึ้น
ทีนี้ Aeon อัพเกรดขึ้นอีกด้วย NVFP4 + Dflash + Uncensor
AEON-7 ปล่อย Orinth 1.0 Ultimate Uncensored — โมเดล AI โค้ดดิ้งเทพ 35B ที่รันบน DGX Spark ได้ เต็มสปีด DFlash ~740 tok/s
ÆON FORGE (@SpaceTimeViking) ประกาศเปิดตัว Orinth 1.0 AEON ULTIMATE UNCENSORED ซึ่งเป็นเวอร์ชัน abliterated/uncensored ของโมเดล Ornith-1.0-35B จาก DeepReinforce AI ที่มีความสามารถด้าน coding สูงระดับต้นแถว
สเปกเด่นๆ:
เป็น MoE 35B parameters (~3B active ต่อ token) — init มาจาก Qwen3.6-35B-A3B
40 layers (30 GatedDeltaNet linear-attention + 10 full-attention), 256 routed experts
มาพร้อม vision tower รองรับ multimodal, 256K context
คะแนน Terminal-Bench 2.1 = 64.2, SWE-bench Verified = 75.6
ผ่านการ abliterate จน 0% refusals โดยที่ coding capability ไม่ลดลงเลย
亮点ใหญ่: มี NVFP4 quantization variant (~23.7GB) สำหรับ DGX Spark / Blackwell architecture โดยเฉพาะ — เล็กพอที่จะรันบน GB10 (128GB unified memory) และเมื่อจับคู่กับ DFlash speculative decoding สามารถทำความเร็วได้ถึง ~740 tok/s aggregate ที่ concurrency 64
AEON-7 ทำ container พิเศษมาให้: ghcr.io/aeon-7/aeon-vllm-ultimate:latest (vLLM 0.23.0 built from source for sm_121a) พร้อม production deployment guide บน GitHub
ใครที่มี DGX Spark อยู่แล้ว นี่คือโมเดลที่น่าสนใจมากสำหรับลองเทียบกับ Qwen3.6-35B-A3B-heretic ที่มีอยู่เดิม โดยเฉพาะเรื่อง coding capability + ไม่มี filter
Credit: @SpaceTimeViking · โมเดลบน HuggingFace: AEON-7/Ornith-1.0-35B-AEON-Ultimate-Uncensored
