Atlas / Models / Rekaai / Reka: Flash 3

Reka: Flash 3✓ Catalog verified

rekaai/reka-flash-3

Reka Flash 3 is a general-purpose, instruction-tuned large language model with 21 billion parameters, developed by Reka. It excels at general chat, coding tasks, instruction-following, and function calling. Featuring a 32K context length and optimized through reinforcement learning (RLOO), it provides competitive performance comparable to proprietary models within a smaller parameter footprint. Ideal for low-latency, local, or on-device deployments, Reka Flash 3 is compact, supports efficient quantization (down to 11GB at 4-bit precision), and employs explicit reasoning tags ("<reasoning>") to indicate its internal thought process. Reka Flash 3 is primarily an English model with limited multilingual understanding capabilities. The model weights are released under the Apache 2.0 license.

Input price
$0.10 /1M
Output price
$0.20 /1M
Context
66K
Modalities
text
Released
Mar 12, 2025
Tool calling
Atlas signal
80/100
01

Overview

Reka Flash 3 is a general-purpose, instruction-tuned large language model with 21 billion parameters, developed by Reka. It excels at general chat, coding tasks, instruction-following, and function calling. Featuring a 32K context length and optimized through reinforcement learning (RLOO), it provides competitive performance comparable to proprietary models within a smaller parameter footprint. Ideal for low-latency, local, or on-device deployments, Reka Flash 3 is compact, supports efficient quantization (down to 11GB at 4-bit precision), and employs explicit reasoning tags ("<reasoning>") to indicate its internal thought process. Reka Flash 3 is primarily an English model with limited multilingual understanding capabilities. The model weights are released under the Apache 2.0 license.

Access: available through the official Rekaai API. Context window: 66K.

Pricing and metadata from the ModelsAtlas catalog.Last refreshed Aug 5, 2026
02

Specifications

API identifierrekaai/reka-flash-3
ProviderRekaai
Model typeText LLM
Context window66K catalog
Input modalitiestext
Output modalitiestext
ReleasedMar 12, 2025
ModeratedNo
Architecture modalitytext->text
Supported parameters
frequency_penaltyinclude_reasoningmax_tokenspresence_penaltyreasoningseedstoptemperaturetop_ktop_p
Source: provider documentation + catalog feedMethodology →

API defaults

Default parameters
{
  "top_k": null,
  "top_p": null,
  "temperature": null,
  "presence_penalty": null,
  "frequency_penalty": null,
  "repetition_penalty": null
}
03

Pricing

Live pricing components for rekaai/reka-flash-3 as published in the catalog.

$0.10
Input /1M
$0.20
Output /1M
Per request
Per image
Pricing componentRaw unit priceNormalizedUnit
Prompt tokens$1e-7$0.10 / 1MPer input token
Completion tokens$2e-7$0.20 / 1MPer output token
Request feeN/AN/APer request
Image feeN/AN/APer image unit
Web search feeN/AN/APer search request
Source: catalog pricing feedCompare all pricing →Cheapest models →
04

Cost calculator

Estimate monthly spend from your own token volumes.

Scale / Volume

Estimated Total Cost

$0.09Calculated from current list pricing in the ModelsAtlas catalog.
05

Capabilities

CapabilityStatusWhat it means
Visual UnderstandingNot advertisedImage and document analysis support
Audio ProcessingNot advertisedSpeech and voice aligned flows
Tool CallingNot advertisedSupports tools / function calling
Self-HostingNot advertisedDeploy outside managed APIs
Derived from catalog capability tags and supported parameters
06

Capability signals

Directional signals derived from model metadata and capability tags — not official benchmark submissions.

Reka: Flash 3
MMLU Signal (Reasoning)
92signal
Coding Signal (HumanEval proxy)
74signal
Math Signal (GSM8K proxy)
78signal
Science Signal (GPQA proxy)
74signal
Estimated from metadata — not an official benchmark runAll benchmarks →Methodology →
07

Quick start

Call Reka: Flash 3 through an OpenAI-compatible client.

from openai import OpenAI

client = OpenAI(
    base_url="https://openrouter.ai/api/v1",
    api_key="YOUR_API_KEY"
)

response = client.chat.completions.create(
    model="rekaai/reka-flash-3",
    messages=[{"role": "user", "content": "Explain quantum physics."}]
)

print(response.choices[0].message.content)
08

Alternatives

Closest models by context window from a different provider.

09

Sources & attribution

Pricing and metadata are maintained in the ModelsAtlas catalog.

Capability bars use metadata tags — directional estimates onlyHow we source and verify data →
10

Frequently asked questions

How much does Reka: Flash 3 cost?

$0.10 per 1M input tokens and $0.20 per 1M output tokens on the official Rekaai API.

What is the context window of Reka: Flash 3?

Reka: Flash 3 supports up to 66K tokens of context.

Does Reka: Flash 3 support tool / function calling?

Reka: Flash 3 does not advertise first-class tool/function calling. Check vendor docs for the latest capabilities.

How do I access Reka: Flash 3?

Use the official Rekaai API with the model id `rekaai/reka-flash-3`. See the Quick start section above for code examples.