Cogito v2.1 671B
Deep Cogito · released Nov 13, 2025
- Type
- Proprietary
- Input
- None held
- Output
- None held
- Cached
- None held
We don't hold a list price for this model yet · hosted only — we hold no downloadable copy
Our take
Written Aug 3, 2026Cogito v2.1 is a proprietary text-only reasoning model from Deep Cogito with a 128,000-token request limit and flat pricing across its two hosted endpoints. No benchmark scores or parameter count have been published, so its capabilities must be judged through direct use rather than leaderboards.
Pick this for long-document text work where a flat rate with no input-output gap matters, or if you are already on OpenRouter or Together and want a single-vendor proprietary option with known throughput on one endpoint. Skip it if you need image, audio or video input, if you want to shop between providers for better speed or price, or if you rely on published benchmark scores to choose a model.
The case for it
- Flat, predictable pricing with no premium for generation: the same rate applies to input and output on both providers.
- Known throughput of 21 tokens per second on the Together endpoint.
The case against it
- No verified quality scores in our data: chat, reasoning, coding and other task performance are all unverified.
- Parameter count undisclosed by the vendor, making hardware planning and efficiency comparisons impossible.
- Only two tracked offers at identical rates with no price competition, and text-only in a multimodal market.
How good is it?
We hold no score for this model.
We look for every model we track on every board we watch, and none of them has turned up Cogito v2.1 671B — so there is no intelligence, coding, agentic or writing score to show you, not a low one, none.
Or rent it from someone else
Cheapest of 2 live listings. Picked at the widest standard context we hold, within one quantisation slice, so the numbers beside it are a price one host actually charges.
- per 1M tokens
- $1.25 in / $1.25 out
- Context served
- 128K
- Throughput
- Not measured
| Provider | In / out per 1M tokens | Context | Throughput | Trains on prompts | Logs prompts | Zero retention |
|---|---|---|---|---|---|---|
| OpenRouter | $1.25 / $1.25 | 128K | not measured | Unknown | Unknown | Unknown |
| Together AI | $1.25 / $1.25 | 128K | 42 tok/s | No | No | Confirmed |
Across the 2 listings we hold: 1 say they do not train on prompts, 0 say they do and 1 do not say. 1 appear in the zero-retention registry we check; the rest are unknown to us rather than confirmed either way.
What each host's API supports
From the parameter list each endpoint publishes. Streaming is omitted: nothing we hold reports it, for any model.
- ✓
- Supported
- ✗
- Not supported
- Not published
- host gave no parameter list
| Provider | Tool calling | JSON output | Strict schema |
|---|---|---|---|
| OpenRouter | ✗ | ✓ | ✓ |
| Together AI | ✗ | ✓ | ✓ |
Tool calling: 0 of 2 listings say yes, 2 say no. JSON output: 2 of 2 listings say yes. Strict schema: 2 of 2 listings say yes.
When we formed this view
Dates behind this page
Prices last checked 9d ago
What we do not know about this model yet
- No board we watch has turned up a score, so we hold no quality figures at all.
- Nothing we hold says whether an endpoint streams, so we do not show it either way.
- 1 of 2 listings do not say whether they train on prompts.
- We don't hold a list price for this model yet — the gap is ours, not the lab's.
- We hold no cached-input rate for any of its listings.
Licence and identifiers
What the licence allowsWe hold no licence record for this model. Inside are the identifiers you need to pull it — its Hugging Face repo where we have one, our slug and a machine-readable card.
Licence
Commercial API terms. We hold no licence record for this model, so there is nothing to summarise here.
Identifiers
- Modality record
- text->text
- Catalogue slug
- deepcogito-cogito-v2-1-671b