najdarenaRequest evaluation
Model catalog

Explore models

Browse models evaluated on Arabic and Saudi tasks. Inspect a model’s results and specifications, or use Leaderboards to compare task rankings.

Historical evaluation dates and publication dates have not been verified. Read the evidence and scoring policy →

Operational metrics

Partial execution-time samples are available under Inference. Controlled speed, token throughput, and cost are not established and are not used to rank these models.