How to choose between models for the 512GB max m3ultra
Ouverte
- Langage dominant
- C
- Étoiles
- 22.4k
- Forks
- 2.1k
- Merge moyen
- 1 j 3 h
- PR mergées (30 j)
- 4
Description
I can see there is different prefill and tokens/sec but how much smarter is pro-imatrix vs q4-imatrix? also how about q2-q4-imatrix? any insight is much appreciated.
Guide de contribution
Ouvrir le guide de contribution
Évaluation
Cette issue n'a pas encore été évaluée.