← Back to leaderboard
Head-to-head
/home/thomaskyn/models/Qwen3.5-4b/Qwen3.5-4B-Q4_K_M.gguf vs gpt-oss:latest
gpt-oss:latest wins+33.8 pts (+58%)on the balanced composite.
Category-by-category
Higher bar wins each row.
/home/thomaskyn/models/Qwen3.5-4b/Qwen3.5-4B-Q4_K_M.gguf
Category
gpt-oss:latest
-50.050.0
Code
100.0+50.0
-40.040.0
Reason
80.0+40.0
87.5
Tool Use
87.5
-12.587.5
RAG
100.0+12.5
-58.829.2
Speed
88.1+58.8
/home/thomaskyn/models/Qwen3.5-4b/Qwen3.5-4B-Q4_K_M.gguf
- Provider
- local
- Released
- Context
- 0K
- Lab
- Community
gpt-oss:latest
- Provider
- local
- Released
- Context
- 0K
- Lab
- Community