PipelineScore
← Back to leaderboard
Head-to-head

/home/thomaskyn/models/Qwen3.5-4b/Qwen3.5-4B-Q4_K_M.gguf vs gpt-oss:latest

gpt-oss:latest wins+33.8 pts (+58%)on the balanced composite.

Category-by-category

Higher bar wins each row.

/home/thomaskyn/models/Qwen3.5-4b/Qwen3.5-4B-Q4_K_M.gguf
Category
gpt-oss:latest
-50.050.0
Code
100.0+50.0
-40.040.0
Reason
80.0+40.0
87.5
Tool Use
87.5
-12.587.5
RAG
100.0+12.5
-58.829.2
Speed
88.1+58.8
/home/thomaskyn/models/Qwen3.5-4b/Qwen3.5-4B-Q4_K_M.gguf
Provider
local
Released
Context
0K
Lab
Community
gpt-oss:latest
Provider
local
Released
Context
0K
Lab
Community
/home/thomaskyn/models/Qwen3.5-4b/Qwen3.5-4B-Q4_K_M.gguf vs gpt-oss:latest · PipelineScore