UPDATE: I will be updating this section as often as I can so you don't have to read the whole thread to find the model results, I will be using humaneval and LCB for the evaluations, I am not a pro if a result does not look good for you please bench it and share your results. My hardware 9070XT + 3060TI(8GB) = 24GB VRAM, 32 RAM, 9800X3D CPU. All models can be downloaded using LM Studio, just paste the model name and download the Quant that fits your hardware, I added the links to huggingface as well.
Ok guys I will stop testing for now unless something or someone finds a good model to test, for now 27B has the top intelligence and if you want speed MOE.
🏆 Visual Leaderboard
| 🏆 Rank |
Model |
Quant |
TPS |
LCB+HEval |
Notes |
| 🥇 |
DavidAU 711 Fable Fusion Heretic NEO MAX MTP |
Q4_K_S |
41.91 |
84.47 |
🏆 Current leader. Tested with 80k context. Best overall benchmark. (HF) |
| 🥈 |
DavidAU 711 Fable Fusion Heretic NEO MAX MTP |
Q4_K_M |
42.93 |
84.35 |
Previous leader. Fastest of the top 3. (HF) |
| 🥉 |
DavidAU 711 Fable Fusion Heretic NEO MAX |
Q4_K_M |
24.65 |
83.81 |
Original non-MTP release. Slow but smart. (HF) |
| 4 |
Architect Polaris2 F451 Heretic |
Q4_K_M |
26.88 |
81.16 |
Heretic version of DavidAU 711. (HF) |
| 5 |
Tess-4-27B |
Q4_K_M |
26.94 |
80.94 |
Excellent all-around coding model. (HF) |
| 6 |
Kwaipilot_KAT-Coder-V2.5-Dev-APEX-MTP-I-Compact |
APEX |
117.59 |
74.93 |
Fastest model tested so far. (HF) |
| 7 |
Genesis Hermes V3 APEX Compact |
APEX |
100.30 |
70.57 |
Excellent speed with solid coding performance. (HF) |
| 8 |
REAM-192 Heretic APEX IBalanced |
Q5_K_M |
81.38 |
69.73 |
Balanced Q5 model. (HF) |
| 9 |
Defiant 9B Heretic MAX MTP |
Q8_0 |
90.79 |
69.56 |
Strongest 9B model tested. (HF) |
| 10 |
Genesis Hermes V5 APEX Compact |
APEX |
100.45 |
65.08 |
High-throughput compact model. (HF) |
| 11 |
Genesis Hermes V4 MTP |
MTP |
103.14 |
64.78 |
High-speed MTP variant. (HF) |
| 12 |
Qwen3.6-27B-A3B Coder |
— |
105.06 |
64.49 |
Official coder baseline. (HF) |
| 13 |
Genesis Hermes V6 |
— |
105.08 |
59.70 |
Latest Hermes release. (HF) |
🏅 Top 3
| 🥇 |
🥈 |
🥉 |
DavidAU 711 MTP Q4_K_S |
DavidAU 711 MTP Q4_K_M |
DavidAU 711 Q4_K_M |
| 84.47 |
84.35 |
83.81 |
| 🚀 41.91 TPS |
🚀 42.93 TPS |
🚀 24.65 TPS |
📊 Performance Snapshot
LCB+HEval
84.47 ████████████████████████████████████████ 🥇 DavidAU 711 Q4_K_S MTP
84.35 ██████████████████████████████████████▉ 🥈 DavidAU 711 Q4_K_M MTP
83.81 █████████████████████████████████████▋ 🥉 DavidAU 711 Q4_K_M
81.16 ██████████████████████████████████ Architect Polaris2
80.94 █████████████████████████████████ Tess-4-27B
74.93 ██████████████████████████████ Kwaipilot KAT-Coder V2.5 Dev
70.57 ███████████████████████████ Genesis Hermes V3
69.73 ██████████████████████████▉ REAM-192
69.56 ██████████████████████████▊ Defiant 9B
65.08 ████████████████████████ Genesis Hermes V5
64.78 ███████████████████████▉ Genesis Hermes V4
64.49 ███████████████████████▊ Qwen3.6-27B-A3B Coder
59.70 █████████████████████ Genesis Hermes V6
UPDATE: I will be updating this section as often as I can so you don't have to read the whole thread to find the model results, I will be using humaneval and LCB for the evaluations, I am not a pro if a result does not look good for you please bench it and share your results. My hardware 9070XT + 3060TI(8GB) = 24GB VRAM, 32 RAM, 9800X3D CPU. All models can be downloaded using LM Studio, just paste the model name and download the Quant that fits your hardware, I added the links to huggingface as well.
Ok guys I will stop testing for now unless something or someone finds a good model to test, for now 27B has the top intelligence and if you want speed MOE.
🏆 Visual Leaderboard
🏅 Top 3
Q4_K_S
Q4_K_M
Q4_K_M
📊 Performance Snapshot