I think smaller models will keep getting better. For privacy and economic reasons, it makes sense to do as much AI on-device as possible.
I think smaller models will keep getting better. For privacy and economic reasons, it makes sense to do as much AI on-device as possible.
This was a MacBook, but since you're just saying "Mac", I don't know what you're referring to.
I think that the MacBook couldn't cope with the heat generated so it went to its knees. I've read about it on Reddit as well.
EDIT: maybe the YouTube demo was using 2 bit quantization?
I suspect OBS can crash it all by itself.
The big thing is just don’t lock the models into memory on lower RAM systems. That gets you into trouble when some of the unified RAM is needed for something else and you can soft lock the system. The heat is not an issue though in my experience.
https://github.com/ml-explore/mlx-examples
This is where all of the action is at right now.
You can also use LM Studio: https://lmstudio.ai/
I struggled to get it working on my machine, it kept corrupting the models as it downloaded them.