Home/Models/Stable Beluga 2

Stable Beluga 2

by Unknown · Released Jan 2024

60.1
avg score
Rank #88
Compare
Better than 72% of all models
Context
N/A
Input $/1M
n/a
Output $/1M
n/a
Type
text
License
Proprietary
Benchmarks
13 tested
Data as of
About

Tested on 13 benchmarks · BenchGecko score 60.1. Top scores: ARC AI2 (81.5%), HellaSwag (78.8%), LAMBADA (71.3%).

Capabilities
reasoning
38.9
#103 globally
math
37.0
#165 globally
knowledge
55.9
#86 globally
general
41.3
#63 globally
language
37.9
#137 globally
Benchmark Scores
Compare All
Tested on 13 benchmarks · Ranked across 5 categories
Score Distribution (all 312 models)
0255075100
▲ You are here
BBH

BIG-Bench Hard. 23 challenging tasks from BIG-Bench where prior language models fell below average human performance.

59.1·
MUSR

HuggingFace MuSR (Multi-Step Reasoning). Tests multi-hop reasoning requiring chaining multiple facts together.

18.6·
GSM8K

Grade school math word problems. 8,500 problems testing multi-step arithmetic reasoning. A foundational math benchmark.

69.6·
MATH Level 5

HuggingFace evaluation of MATH Level 5 problems. Competition math requiring advanced reasoning and proof construction.

4.4·
ARC AI2

AI2 Reasoning Challenge. Grade-school science questions requiring multi-step reasoning. Easy and Challenge sets test different difficulty levels.

81.5·
HellaSwag

Sentence completion requiring commonsense reasoning about physical and social situations. Tests real-world understanding.

78.8·
LAMBADA

Language modeling benchmark testing ability to predict the last word of passages requiring long-range context understanding.

71.3·
Excellent (85+) Good (70-85) Average (50-70) Below (<50)
Links
Documentation
Community
BenchGecko API
stable-beluga-2
Specifications
  • Typetext
  • ContextN/A
  • ReleasedJan 2024
  • LicenseProprietary
  • Statusbenchmark-only
Available On
Unknownn/a
Share & Export
Tweet
Stable Beluga 2 is a proprietary text AI model by Unknown, released in January 2024. It has an average benchmark score of 60.1.

Key facts · as of 2026-04-09

  • Stable Beluga 2 by Unknown. BenchGecko score 60.1, rank 87 of 312 scored models (normalized average of public benchmark scores).
  • List price n/a input · n/a output per 1M tokens (as of 2026-04-09).

How to cite · data as of 2026-04-09

Stable Beluga 2 · benchmarks, pricing and providers. BenchGecko, data as of 2026-04-09. https://benchgecko.ai/model/stable-beluga-2

Credit "Source: BenchGecko" with a link. Prices per provider and Gecko Tests are BenchGecko data (CC BY 4.0); benchmark scores keep their original source, listed in the JSON. JSON · llms.txt · MCP