Whats your Qwen vs Gemma take on local hardware Marky? I find Gemma WAY faster on old CPU based compute where I run open models on an old laptop, so very anecdotal and not scientific but Gemma is actually usable on an old Lenovo in ollama, Qwen had a literal mental breakdown just trying to answer "what model variant are you?" as a first prompt. It went into a mental meltdown about how to respond that was so bizarre I actually saved it as a text file while it looped over its identity crisis about how to respond to me for over 15 minutes and I had to actually break it and stop it. Gemma responded with its model variant in about 11 seconds.
VERY unscientific but using it is what really matters. I wanted to use Qwen but it just won't run on moldy old hardware and Gemma does ok there.
So I am curious what your benches say about them compared to each other on accuracy?
RE: Qwen 3.8 27B Released - My observations