Pokee-Isaac 28B

Pokee-Isaac 28B is a reasoning model from Pokee AI in the Pokee-Isaac family. 3 benchmarks count toward its score, in 1 category.

partialShows if the model has enough results for an index.
IndexOverall score. 50 is the middle.Unranked
CoverageShare of the index weight with results.20%
SpeedOutput tokens per second.
Input / 1MUS dollars per 1M input tokens.$0.15
Output / 1MUS dollars per 1M output tokens.$1
ContextMaximum tokens in one request.10M
EloLMArena rating and rank.N/A

50 is the middle of the board. The range shows the doubt in the index.

CapabilitiesScore per category. 50 is the middle.

50 is the middle
AgenticMulti-step tasks with tools.
56.6
CodingCode writing and repair.
N/A
ReasoningLogic problems and puzzles.
N/A
MultimodalTasks with images and text.
N/A
KnowledgeFacts and expert knowledge.
N/A
MultilingualTasks in many languages.
N/A
InstructionTasks with strict rules in the prompt.
N/A
MathMath problems.
N/A

Results

3 counted
BenchmarkThe test name.CategoryThe capability that the test measures.ResultThe score from the publisher.IndexThis result on the index scale.RunThe settings of the run.DateDate of the result.Published byThe source of the result.
PinchBenchAgentic95.7%Kilo Code
MCP-Atlas mean claim coverageAgentic74.6%Anthropic
Berkeley Function Calling Leaderboard v4Agentic70.9%55.9Arcee AI
τ³-Bench Tool-Agent-User EvaluationAgentic66.2%51.4Sierra Research
Terminal-Bench 2.1 (provider run)Agentic65.1%62.5DeepSeek-AI
Terminal-Bench 2.1 (provider run)Agentic65.1%62.5DeepSeek-AI
MRCRv2Reasoning60.7%OpenAI

3 benchmarks count, from 4 of 7 results. A grey row does not count. Too few models took that benchmark.

Sources

BenchLM benchmark aggregationUsed with attribution; per-benchmark results credited to their original authors