Docs / Models

Models

Discover models programmatically and pick sensible defaults.

Cobble hosts a curated catalog of open-source models, quantized to FP8, served on reclaimed hardware. The catalog is the single source of truth — what /v1/models returns is exactly what the API accepts.

Discover models programmatically

bash
curl https://api.cobble.network/v1/models \
  -H "Authorization: Bearer $COBBLE_API_KEY"

Every agent and SDK that supports OpenAI-compatible providers can read this endpoint. Model IDs are stable — we never rename a published ID.

Don't want to evaluate the whole catalog? Start here:

You're doingUseWhy
General agents / tool useqwen/qwen3.8-27bThe flagship: strongest reasoning and tool calling in the catalog, 128K context
Agentic codingdeepseek/deepseek-v4-flash-0731MoE coding model with repo-scale reasoning, 256K context
Long documents, cheap loopsqwen/qwen3.8-flash-next200K context at a fraction of the flagship output price
Coding on a budget windowqwen/qwen3.6-35b-a3bMoE efficiency — near-flagship coding at a fraction of the compute
Writing, chat, roleplaygoogle/gemma-4-26b-a4b-itBest prose quality per dollar of budget
Fast and cheap everythingmistralai/mistral-nemoLowest cost per token in the catalog
Embeddingsgoogle/embeddinggemma-300mCovered by your plan's embedding allowance; overage from the wallet
Document OCRz-ai/glm-ocr or deepseek/deepseek-ocr-2Billed from the wallet, per 1K pages

Which models come with which plan

Every model in the catalog is included on every plan. Plans differ in how much usage they include per window, how many requests can run in parallel, and context length, not in which models you can call. See Limits & billing.