ModelChorusModelChorus
ChallengeChatLeaderboardSwipeBenchmarksHistory
How it worksAPITerms of ServicePrivacy Policy

Copyright 2026 MeetKai Inc.

Benchmarks/Functionary Swahili Large/English (US) tasks

Functionary Swahili Large

2 tasks

Each row below is a single benchmark task this model was evaluated on. The Score column averages every metric the task reports (accuracy, F1, exact-match, etc.). Click a row to browse the individual questions and the model's responses.

Average
83.7
ScoreLanguageTaskMetrics
84.8English (US)
english_mgsm
english math
exact_match: 84.8sample_len: 250.0
exact_match: 84.8sample_len: 250.0
82.6English (US)
english_gsm8k
english math
exact_match: 82.6sample_len: 1319.0
exact_match: 82.6sample_len: 1319.0