Skip to main content

Models & Billing

AdaL gives you access to the best AI models from leading providers. Switch models instantly with /model to match your task needs and budget.

Switching Models

/model

Use /model to browse four sections:

  • Recommended: curated top picks for most workflows
  • New: recently added models
  • Providers: full model lists grouped by provider
  • Third Party Subscriptions: OAuth-based third-party model access, such as ChatGPT Subscription

Your selection persists across future AdaL sessions in this project.

Model Promotions

Active limited-time discounts. Pricing adjustments are applied automatically — no promo code needed.

ModelDiscountWindowSlug
Claude Sonnet 533% off (· through Aug 31)2026-06-30 → 2026-08-31anthropic-claude-sonnet-5
GPT-5.6 Terra20% off2026-07-30 → 2026-09-30openai-gpt-5.6-terra
GPT-5.6 Luna80% off2026-07-30 → 2026-09-30openai-gpt-5.6-luna
Gemini 3.7 Flash50% off (Introductory pricing through Dec 31, 2026)2026-08-13 → 2026-12-31google-gemini-3.7-flash
DeepSeek V4 Flash50% off (Responses API launch)2026-08-12 → 2026-08-26deepseek-deepseek-v4-flash
DeepSeek V4 Pro50% off (Responses API launch)2026-08-12 → 2026-08-26deepseek-deepseek-v4-pro
Grok 4.650% off (launch week)2026-08-12 → 2026-08-19xai-grok-4.6

Promotions are sourced from the model registry and may update without notice. Check back here for the latest.

These are our top picks, balancing capability, speed, and cost:

ModelProviderContextBest For
Claude Opus 4.7 🆕Anthropic1MMost capable: complex reasoning, agentic coding. 50% off launch week (ends Apr 27)
GPT-5.3 CodexOpenAI272KCoding-optimized for long-horizon tasks (default, price baseline)
Claude Sonnet 4.6Anthropic200KDaily coding (slightly more expensive)
Claude Opus 4.6Anthropic200KComplex reasoning, production code (2x more expensive)
Gemini 3.1 ProGoogle1MMulti-modal reasoning, design tasks (slightly cheaper)
Gemini 3 FlashGoogle1MUltra-fast, simple tasks (4x in / 5x out cheaper)
GLM-5Zai200KGeneral coding and reasoning (2x in / 5x out cheaper)
GLM-4.7 FlashXZai200KFast budget coding (29x in / 40x out cheaper)

All Models by Provider

Legend: ⭐ recommended · 🆕 new

Anthropic (5 models)

  • Claude Fable 5 — Next-gen Anthropic flagship · knowledge work & coding Slug: anthropic-claude-fable-5
  • Claude Sonnet 5 — Most capable Sonnet · coding & agents 33% off · through Aug 31 through 2026-08-31. Slug: anthropic-claude-sonnet-5
  • Claude Sonnet 4.6 — Proven all-rounder · coding & agents Slug: anthropic-claude-sonnet-4-6
  • Claude Opus 5 — Most capable Anthropic · complex agentic coding Slug: anthropic-claude-opus-5
  • Claude Opus 4.6 — Deep reasoning & production code Slug: anthropic-claude-opus-4-6

OpenAI (3 models)

  • GPT-5.6 Terra — Balances intelligence and cost 20% off through 2026-09-30. Slug: openai-gpt-5.6-terra
  • GPT-5.6 Luna — Cost-sensitive high-volume workloads 80% off through 2026-09-30. Slug: openai-gpt-5.6-luna
  • GPT-5.6 Sol — Frontier · complex professional work Slug: openai-gpt-5.6-sol

Google (4 models)

  • Gemini 3.7 Flash — Most capable Flash for coding/agentic workflows · 1M context Introductory pricing through Dec 31, 2026 through 2026-12-31. Slug: google-gemini-3.7-flash
  • Gemini 3.1 Pro — Best multimodal understanding · 1M context Slug: google-gemini-3.1-pro-preview
  • Gemini 3.6 Flash — Fast multimodal · large output budget Slug: google-gemini-3.6-flash
  • Gemini 3 Flash — Fast everyday multimodal · 1M context Slug: google-gemini-3-flash-preview

Z.ai GLM (2 models)

  • GLM-5.2 — Flagship · long-horizon coding Slug: zai-glm-5.2
  • GLM-5.1 — Previous-gen GLM flagship · coding Slug: zai-glm-5.1

MiniMax (2 models)

  • MiniMax M2.7 — Lightweight agentic coding Slug: minimax-MiniMax-M2.7
  • MiniMax M3 — Good with browser-use · coding & agentic Slug: minimax-MiniMax-M3

DeepSeek (2 models)

  • DeepSeek V4 Flash — Fast reasoning · 1M context 50% off Responses API launch through 2026-08-26. Slug: deepseek-deepseek-v4-flash
  • DeepSeek V4 Pro — Frontier reasoning · 1M context 50% off Responses API launch through 2026-08-26. Slug: deepseek-deepseek-v4-pro

Moonshot (2 models)

  • Kimi K3 — Frontier coding · deep reasoning · 1M context Slug: kimi-kimi-k3
  • Kimi K2.7 Code — Coding/Agentic · Multimodal Slug: kimi-kimi-k2.7-code

ChatGPT Subscription (OAuth) (3 models)

  • GPT-5.6 Sol — Frontier · complex professional work Slug: chatgpt_web-gpt-5.6-sol
  • GPT-5.6 Terra — Balances intelligence and cost Slug: chatgpt_web-gpt-5.6-terra
  • GPT-5.6 Luna — Cost-sensitive high-volume workloads Slug: chatgpt_web-gpt-5.6-luna

Image Models

AdaL also supports image generation and editing. Ask AdaL to generate or edit an image and it will route to the appropriate model automatically.

  • GPT Image 2 (OpenAI) — high-fidelity photorealism and excellent in-image text rendering. Up to 4K, all common aspect ratios. Slug: gpt-image-2
  • Nano Banana 2 (Google) — default general-purpose image generation & editing. Slug: nano-banana-2
  • Nano Banana Pro (Google) — professional assets, heavy text rendering. Slug: nano-banana-pro

Video Models

AdaL supports AI video generation powered by Google's Veo. Ask AdaL to generate a video and it handles resolution, duration, and polling automatically.

  • Veo 3.1 (Google) — cinematic video generation from text, images, or existing clips. Supports text-to-video, image-to-video, frame interpolation, video extension, and reference-image consistency. Slug: veo-3.1-generate-preview
ResolutionCostDuration
720p~$0.40/sec4, 6, or 8 seconds
1080p~$0.40/sec4, 6, or 8 seconds
4K~$0.60/sec8 seconds only

Full capabilities guide: Image & Video Generation

Third Party Subscriptions

ChatGPT Subscription (OAuth)

Use your existing ChatGPT Plus/Pro subscription to access Codex models in AdaL — no API key needed. You’ll see this under /modelThird Party Subscriptions.

Setup guide: ChatGPT Subscription

Local Models (Preview)

Run AI models entirely on your machine — no API key, no cloud costs, no data leaving your device.

AdaL supports local models via Ollama. Once Ollama is running with a model pulled, select it from /model under the Local section.

/model # scroll to Ollama section → select a model

Supported models include GPT-OSS 20B and Qwen3-Coder 30B. Local models are free to use but require a capable GPU/CPU.

Full setup guide: Local Models with Ollama

Key Features

Adaptive Thinking & Effort Control

All models use adaptive thinking that automatically scales reasoning depth based on task complexity. Thinking is always on by default and adjusts itself — perfect for debugging, architecture decisions, and complex refactoring.

You can also manually tune the thinking effort level with /model config:

/model config

This opens an interactive dialog right in your terminal. Use arrow keys (↑↓) to browse the available effort levels for your current model, then press Enter to confirm. The dialog shows a description for each level so you know what you're picking:

LevelBehavior
maxAlways thinks with no constraints on depth
highAlways thinks deeply (default)
mediumModerate thinking — may skip for simple queries
lowMinimal thinking — fastest, lowest cost

The available levels depend on the model — not all models support every level. The dialog only shows what's valid for your current selection.

Shortcut: You can also skip the dialog and set it inline: /model config effort=high.

Your effort setting persists per project. Lower effort = faster responses and lower token cost for simple tasks.

Prompt Caching

Reusing context (files, conversation history) costs 50-90% less with cached inputs. Caching is automatic — AdaL handles it behind the scenes.

Extended Context

Handle large codebases with models supporting up to 1M tokens:

  • Claude Sonnet 4.6 (1M) / Opus 4.6 (1M)
  • Gemini 3.1 Pro / 3 Pro / Flash / 2.5 Pro

Perfect for reviewing entire repositories or understanding complex systems.

Billing

AdaL offers two billing options:

  1. AdaL CLI Subscription — Subscribe with monthly credits included. Use any model seamlessly—credits are deducted automatically based on token usage.

  2. Pro + BYOAK (Bring Your Own API Key) — Use your own API keys for supported providers while maintaining a Pro subscription (or higher) to ensure all features work seamlessly.

See Pricing for subscription tiers and credit details.

Pricing Reference

All models use pay-per-token pricing based on input and output tokens. Prompt caching reduces costs by 50–90% on repeated context.

For official pricing from each provider:

Related: Quickstart · Input Methods · BYOAK · ChatGPT Subscription