(1/2) š¦ Buckle up and ready for a wild llama ride with 70B Llama-2 on a single MacBook š» š¤Æ Now 70B Llama-2 can be run smoothly on an 64G M2 max with 4bit quantization. š Here is a step-by-step guide: mlc.ai/mlc-llm/docs/g⦠š How about the performance? It's
![]()



