Relace
·Released on Dec 8, 2025

Relace Search API Benchmarks, Pricing & Provider Data

Compare Relace Search with another model

Choose a model to open its comparison page.

Share on X
LLM

The relace-search model uses 4-12 `view_file` and `grep` tools in parallel to explore a codebase and return relevant files to the user request. In contrast to RAG, relace-search performs agentic...

Specifications

Input and output token limits for this model, plus how it ranks on long-context understanding.

INPUT
256Ktokens
307.2 pages of text
OUTPUT
128Ktokens
8K128K1M4M
256K

Features

Technical Details

Input
Output
Released
Dec 2025
Documentation
Tokenizer
Other
Architecture
text->text
Moderated
No
Supported parameters
max_tokensresponse_formatseedstoptemperaturetool_choicetoolstop_p

OpenRouter endpoints

1 endpoints

Third-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.

Provider endpointInputOutput1d uptime30m latency30m throughputContext / output
Relace
relace/bf16
$1/M$3/M100%undefined tokens / undefined tokens

Alternatives & Similar Models

Also known as

relace/relace-search

Rankings are based on community-submitted tests and periodic health probes. Advisory only, not official data.Standard benchmark data may include BenchLM and other public sources.

Build with LMSpeed Data

Free

Free public API for LLM pricing, benchmarks & provider data

  • Real-time pricing & availability
  • Speed & latency benchmarks
  • 300+ models, 600+ providers
  • 1,000 requests/day
View API Documentation