(1/2) ๐ฆ Buckle up and ready for a wild llama ride with 70B Llama-2 on a single MacBook ๐ป ๐คฏ Now 70B Llama-2 can be run smoothly on an 64G M2 max with 4bit quantization. ๐ Here is a step-by-step guide: mlc.ai/mlc-llm/docs/gโฆ ๐ How about the performance? It's
00:00
