Skip to content
Dashboard

Embed 5 Fast

Cohere's lightweight embedding model optimized for low latency and high throughput in interactive search, agent loops, and high-volume retrieval. Embed 5 Fast supports text, images, and mixed content, over 100 languages, and a 128,000-token context window, with output dimensions from 256 to 2,048. It shares an embedding space with Embed 5 Pro. Your use is subject to Cohere's Terms & Privacy Policies.

Input price
Input $0.08, Per 1M tokens
import { embed } from 'ai';
const result = await embed({
model: 'cohere/embed-v5.0-fast',
value: 'Sunny day at the beach',
})
Read docs

Copy link to headingProviders

Route requests across multiple providers. Copy a provider slug to set your preference. Visit the docs for more info. Using a provider means you agree to their terms, listed under Legal.

Checking availability for your team
Provider
Context
Input
Capabilities
ZDR
No Training
Free AI Gateway Credit
Release Date
128K
$0.08/M
09/30/2026

Copy link to headingMore models by Cohere

Model
Context
Latency
Throughput
Input
Output
Cache
Search
Capabilities
Providers
ZDR
No Training
Free AI Gateway Credit
Release Date
128K
$0.12/M
—
cohere logo
09/30/2026
32K
$2/K
—
cohere logo
12/11/2025
32K
$2.50/K
—
cohere logo
12/11/2025
128K
$0.12/M
—
bedrock logo
cohere logo
04/15/2025
256K0.4 s66 tps
$2.50/M
$10/M
—
cohere logo
03/13/2025
4K
$2/K
—
bedrock logo
12/02/2024