A new ML compiler to run 70B+ LLMs on consumer GPUs with <1% accuracy loss

4 points | by marcelbuilds 15 hours ago

3 comments