Artificialanalysis.ai ranks Gemma4 above Qwen3.6 27b in SciCode benchmark results; user questions potential benchmarking issue.
Read the original at old.reddit.com→Just came across this coding benchmark: SciCode Artificialanalysis.ai reports a ranking which contradicts the feeling we've towards those models in real life coding. Is Gemma 4 really that good, or a benchmarking...
Original headline: "How come artificialanalysis.ai ranks Gemma4 above Qwen3.6 27b in SciCode"
Coverage timeline
- Aug 6, 13:21 UTC r/LocalLLaMA lead source How come artificialanalysis.ai ranks Gemma4 above Qwen3.6 27b in SciCode