Skip to content
Build Perch
Browse documentation

Choosing a model

Short answer: The model picker is one searchable, sortable list: every connected provider’s models together, each row showing its context window, its real per-token rate, and whether it supports tools. Each chat is pinned to one model but you can switch at any point without losing history.

What is the model picker?

Every chat picks its model when it’s created, from a single list that combines whatever providers you’ve connected: OpenRouter’s live catalog, your local Ollama models, and Ollama Cloud’s hosted models together. Rather than a curated “recommended” shortlist, the list is fully sortable by name, context window, price, or tool support, and searchable by name or id. Your starred favorites and the 10 most recently added OpenRouter models sit at the top as discovery aids; everything else groups by provider underneath.

Media models (image, speech, music, video generators) aren’t part of this same chat-model list. They’re picked separately, per category, from the relevant row in a chat’s configuration, because only OpenRouter serves media generation at all.

How do I pick or switch a model?

  1. Open Connect AI Model (new chat) or Switch Model (existing chat).
  2. Search by name, or browse the Favorites / Recently Added / per-provider groups. Star a model to pin it to Favorites for next time.
  3. Sort the list by clicking a column header: Model, Context, Rate (in/out), or Tools. Every section reorders together.
  4. Compare rows directly: each shows the model’s context window, its $/1M-token input/output rate (or “Free” / “Subscription” / “See provider” when a clean rate doesn’t apply), and a checkmark if it supports tool calls. Chat models without tool support carry a “No Tools” badge.
  5. Pick a model and connect. Switching an existing chat’s model preserves its history; a banner warns you first if the new model can’t do something the current one can (no tools, no vision, no audio, or a much smaller context window than what’s already in the conversation).

Good to know

  • Every rate links to the model’s page on OpenRouter or Ollama for the full pricing breakdown. There’s no in-app spend total; Build Perch shows per-model rates only, and you pay the provider directly.
  • Media generation is a separate pick, and OpenRouter-only. Image, speech, music, and video models are chosen from their own row in a chat’s configuration, pinned to that one category; neither Ollama provider generates media, so they’re never offered there.
  • Switching models mid-chat is designed for cost-tiering a workflow: run a cheap model for the legwork, then switch the same chat to a frontier model for the final pass. History carries over either way.
  • A model without tools can still receive a mode with tools assigned; the switch banner flags it and offers to fall back that chat to Chat Only automatically on confirm, rather than leaving tools silently broken.
  • Local Ollama models show live capability data (size, quantization, context window, tool/vision support) pulled from Ollama itself, so you’re comparing your actually-installed models on the same terms as the remote catalog.