Technology Innovation Institute, open weights
Falcon-180B
Falcon-180B by Technology Innovation Institute ranks 260th of 354 ranked models on the Noometry Index as of October 2026, with a score of 32.2. Its strongest category is reasoning, where it ranks 269th.
Last verified
Specifications
- Noometry rank
- #260 of 354
- Index score
- 32.2
- Evidence
- Confirmed 16 results
- Provider
Technology Innovation Institute
- Released
- September 6, 2023
- Weights
- Open weights
- Reasoning
- Unknown
- Context window
- —
- Max output
- —
- Input price
- Not listed
- Output price
- Not listed
- Blended price
- Not listed
- Output speed
- Not measured
- Value
- Not ranked
- Knowledge cutoff
- Unknown
Category scores
Each category score combines every public result we have in that category.
- Reasoning 19.1
- Multilingual 25.2
- Instruction Following 53.4
- Writing & Preference 29.1
| Category | Score | Rank | Results |
|---|---|---|---|
| Reasoning | 19.1 | #269 | 1 |
| Multilingual | 25.2 | #286 | 1 |
| Instruction Following | 53.4 | #286 | 1 |
| Writing & Preference | 29.1 | #295 | 3 |
Strengths and weaknesses
Categories where Falcon-180B places highest and lowest among the models ranked in each, with its score against that category's median.
Strongest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Reasoning | 19.1 | −4.5 | #269 of 350, top 77% |
| Instruction Following | 53.4 | −17.9 | #286 of 305, top 94% |
Weakest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Multilingual | 25.2 | −22.2 | #286 of 297, top 97% |
| Writing & Preference | 29.1 | −24.6 | #295 of 312, top 95% |
Closest competitors
The models ranked just above and below Falcon-180B. When scores are this close, price and speed are often the better way to choose.
| Model | Rank | Score | Blended $/M | Speed | |
|---|---|---|---|---|---|
| Laguna M.1 | #256 | 32.5 | — | — | Compare |
| Command R+ | #257 | 32.4 | $4.38 | — | Compare |
| Granite 3.1 8b Instruct | #258 | 32.4 | — | — | Compare |
| Pixtral Large | #259 | 32.2 | $3 | — | Compare |
| Gemini 1.5 Pro (May 2024) | #261 | 32.1 | — | — | Compare |
| Gemma 3 12B | #262 | 32.1 | $0.075 | — | Compare |
| Mistral Large | #263 | 31.9 | $3 | — | Compare |
| Qwen3-4B | #264 | 31.9 | — | — | Compare |
Sponsored placements are available on pages like this one. Advertise on Noometry
Benchmark results
Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.
Reasoning
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Hard Prompts | 1007 | #290 of 297, top 98% | LMArena | 2026-10-08 | |
| Epoch Capabilities Index | 112.13 | #189 of 213, top 89% | Epoch AI | 2023-09-06 | |
| HellaSwag | 89% | #4 of 29, top 14% | Epoch AI | ||
| LAMBADA | 79.8% | Best of 9 | Epoch AI | ||
| PIQA | 84.9% | #5 of 27, top 19% | Epoch AI | ||
| WinoGrande | 87.1% | #5 of 43, top 12% | Epoch AI |
Math
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| GSM8K | 54.4% | #19 of 38, top 50% | Epoch AI |
Knowledge
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| ARC (AI2) Challenge | 67.8% | #19 of 39, top 49% | Epoch AI | ||
| BoolQ | 89% | Best of 23 | Epoch AI | ||
| MMLU | 70.6% | #43 of 81, top 54% | Epoch AI | ||
| OpenBookQA | 64.2% | #10 of 19, top 53% | Epoch AI |
Multilingual
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Non-English | 1000 | #286 of 297, top 97% | LMArena | 2026-10-08 |
Instruction Following
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Instruction Following | 1047 | #283 of 298, top 95% | LMArena | 2026-10-08 |
Writing & Preference
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Text | 1054 | #287 of 297, top 97% | LMArena | 2026-10-08 | |
| LMArena Creative Writing | 1089 | #270 of 295, top 92% | LMArena | 2026-10-08 | |
| LMArena Multi-Turn | 1013 | #286 of 295, top 97% | LMArena | 2026-10-08 |
Compare Falcon-180B
- Falcon-180B vs Pixtral Large
- Falcon-180B vs Gemini 1.5 Pro (May 2024)
- Falcon-180B vs Granite 3.1 8b Instruct
- Falcon-180B vs Gemma 3 12B
- Falcon-180B vs Command R+
- Falcon-180B vs Mistral Large
- Falcon-180B vs GPT-6 Astra
- Falcon-180B vs Claude Fable 5.1
- Falcon-180B vs Gemini 3.8 Flash
- Falcon-180B vs Kimi K3
- Falcon-180B vs Grok 4.6
- Falcon-180B vs Qwen3.8 Max
- Falcon-180B vs GLM-5.3
- Falcon-180B vs Muse Spark 1.3
Other Technology Innovation Institute models
Frequently asked questions
How good is Falcon-180B?
Falcon-180B by Technology Innovation Institute ranks 260th of 354 ranked models on the Noometry Index as of October 2026, with a score of 32.2. Its strongest category is reasoning, where it ranks 269th.
Is Falcon-180B open source?
Yes. Falcon-180B's weights are downloadable; check the license for commercial terms.
What are Falcon-180B's strengths and weaknesses?
Relative to other ranked models, Falcon-180B places best in reasoning, instruction following and lowest in multilingual, writing & preference.
What is Falcon-180B best at?
Its best category is reasoning, where it ranks 269th on Noometry.