评测Gemma - 对比Mistral,Qwen1.5
简要评测谷歌Gemma,对比基线为Qwen1.5,Mistral,与Llama-2.总的来说:代码能力:与Qwen1.5同属第一梯队,超越Mistral,离CodeLlama还有一些距离。数学能力:未超越Qwen1.5,与Mistral并列排第二。文本建模能力:与Mistral有较大差距,但鲁棒性显著超过Lla…
English triage
Chinese technical note: evaluationGemma - 对比Mistral,Qwen1.5
九号 published a Zhihu article relevant to frontier and open model development, evaluation and reliability. The original Chinese excerpt is included below so the feed can preserve the raw source while giving English readers enough context to triage the item.