Hacker news

  • Top
  • New
  • Past
  • Ask
  • Show
  • Jobs

Shapelearn Qwen 3.8 27B (13.1 GB VRAM) (https://byteshape.com)

103 points by syntaxing 2 days ago | 39 comments | View on ycombinator

_ache_ 2 days ago |

From my own test. It's not faster than the unsloth model.

Disclarer: I'm unsing Vulkan on an AMD GC.

syntaxing 1 day ago |

I’m on a strix halo @ GPU-5 with MTP and I get 600 prefill and 30 TG which pushes it into a very usable range. The odd thing is that Dflash2 is really slow for me, like sub 10 TG.

Schlagbohrer 1 day ago |

Absolute treasure of a website with these graphs, thank you for sharing this. Huge help for me to find a faster model (smaller quantization) for my VRAM.

sheo 2 days ago |

npodbielski 1 day ago |

Well I tested it on 7900XTX with the same prompts and their draft model gave me about 30t/s. Their own snippet of code with regular MTP model gave me 60t/s.

Also model with their draft answered incorrectly. With MTP it answered correctly.

Question was: "Does MikroTik CRS312-4C+8XG-RM have combo ports?". The answer is Yes.

kristianp 1 day ago |

What's GPU-5?