kth8/Qwen3.6-27B-insurance-benchmark
Benchmark of Qwen/Qwen3.6-27B against kth8/insurance dataset. Accuracy: 92.0%. Metric Value Correct 46 Incorrect 4 Errors 0 Total samples 50 Total completion tokens 57,899 Raw stats: { "accuracy": 0.92, "correct": 46, "incorrect": 4, "error": 0, "total": 50, "python_tool_calls": 0, "completion_tokens": 57899 }
013
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face