Hy3

No longer available on HF due to storage restrictions - archived here

See Hy3 in action: demonstration videos

Tested with an M3 Ultra 512 GiB using Inferencer app

  • Text inference: ~20 tokens/s @ 1000 tokens ~315 GiB

Screenshot

Downloads last month

-

Downloads are not tracked for this model. How to track
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for inferencerlabs/Hy3-MLX-Q9

Base model

tencent/Hy3
Quantized
(71)
this model