Releasing smolbenchmark: Helps you choose the best model for your hardware!
Most model leaderboards assume a server with powerful GPUs to run models that people daily use.
However, my smolbenchmark is the other column: models that fit in 8GB, ranked by:
- decode speed,
- tokens per joule, and
- heat,
and all of this on your OWN hardware ranging from:
- tablets
- phones
- macs
- jetsons
- raspberry pis
Currently, 13 families on the chart right now, ~1000 configs for the Jetson nano Orin Super 8GB. One device is live measuring:
- tok/s
- tok/J
- ITL
- latency
- power metrics
- thermals and battery
Models that are small enough to actually fit on a device that you own. All the performance benchmarking I did, will be released here for anyone to look at and decide what exact model they would wanna use on their choice of hardware.
Well currently, the Pi, phones, and Mac minis still in the oven, cooking and not filled in yet, but will soon be filled in!
You will now you know which model is BEST for your own hardware with all the raw data available and details reports available to you
https://yuvrajsingh-mist.github.io/smolbenchmark/
(still in heavy development; would love to hear feedback/suggestions on what can be improved!)