Meta-Llama-3-120B-Instruct已经排进Huggingface热门排行Top10,它是一个由"Meta-Llama-3-70B-Instruct"自我合并而成的模型,使用MergeKit工具进行合并的。
来自网友的评价
Llama3-120B 在这些难题上确实展现了比GPT-4更高的智能
GPT-4 -> 不会 Llama3-120B -> 只有在我们质疑量子力学的哥本哈根解释时,让我来解释一下...
https://twitter.com/spectate_or/status/1787308316152242289
让Llama-3-120B解释下面的笑话(实际上是发生的)
https://twitter.com/spectate_or/status/1788031383052374069
llama3-120B 在 bfloat16 格式下表现相当出色
https://twitter.com/_xjdr/status/1787666447612985456
有趣的话题:Meta-Llama3-120B原生的自我合并Llama3以击败GPT4
虽然并不倡导视频中的所有观点
https://twitter.com/GG_Ashbrook/status/1788365679860596957
Llama3-120B版本交流——这玩意儿太聪明了
它不再让我随意摆布。它有自己的主意。
https://twitter.com/erhartford/status/1787050962114207886
Ollama+Llama3-120b
通过ollma使用llama3-120b-Q4-K_M量化版本,48G显存、38G RAM就可以run起来
slices:- sources:- layer_range: [0, 20]model: meta-llama/Meta-Llama-3-70B-Instruct- sources:- layer_range: [10, 30]model: meta-llama/Meta-Llama-3-70B-Instruct- sources:- layer_range: [20, 40]model: meta-llama/Meta-Llama-3-70B-Instruct- sources:- layer_range: [30, 50]model: meta-llama/Meta-Llama-3-70B-Instruct- sources:- layer_range: [40, 60]model: meta-llama/Meta-Llama-3-70B-Instruct- sources:- layer_range: [50, 70]model: meta-llama/Meta-Llama-3-70B-Instruct- sources:- layer_range: [60, 80]model: meta-llama/Meta-Llama-3-70B-Instructmerge_method: passthroughdtype: float16
https://hf-mirror.com/mlabonne/Meta-Llama-3-120B-Instruct
