EdgazeDocsEdgaze Docs
DocumentationAPI Reference
Home
Legal & Trust Center
Core Platform
Terms of ServicePrivacy PolicyCookie PolicyAcceptable Use PolicyPlatform Status and Beta DisclaimerContent Disclaimer
Marketplace & Payments
Payments OverviewMarketplace FeesCreator EarningsPayout SystemRefund PolicyChargeback PolicyPricing Limits
Creator Rules
Creator TermsCreator GuidelinesWorkflow Run PolicyInfrastructure Cost EstimationModel Catalog & PricingCreator Subscription Policy
Trust, Safety & IP
Community GuidelinesFraud and Abuse PolicyDMCA and IP Takedown PolicySecurity and Responsible Disclosure

Platform status

Checking platform status
Legal & Trust Center / Creator Rules

Model Catalog & Pricing

The AI models Edgaze workflows can call, what hosted runs cost per model, and how older model ids are automatically migrated to current ones.

Lists the AI model catalog with hosted per-model rates and explains automatic migration of retired model ids.
Applies to Creators, Customers.

Overview#

Every LLM block in an Edgaze workflow (LLM Chat, LLM Image, Embeddings, AI Conditions) runs against a curated model catalog. This page lists the current catalog, the hosted platform rate for each model, and explains what happens to workflows that reference older models.

The model table is rendered from the same catalog used by the builder and billing engine. Provider prices and capabilities can still change between releases, so confirm the current publish preview before relying on an estimate.

For how per-run estimates are produced, read Infrastructure Cost Estimation. For how estimates combine with creator margin, read Marketplace Fees.

Chat Models#

ModelBest forInput (per 1M tokens)Output (per 1M tokens)
GPT-5.6 TerraRecommendedHigh-quality general workflows$2.70$16.20
GPT-5.6 SolFlagship reasoning$5.40$32.40
GPT-5.6 LunaFast, high-volume steps$1.08$6.48
Kimi K2.6Agentic open model$1.03$4.32
DeepSeek V4 ProDeep reasoning value$1.88$3.76
Claude Sonnet 5Frontier coding + agents$3.24$16.20
Claude Opus 4.8Premium reasoning$5.40$27.00
Gemini 3.5 FlashHigh-quality multimodal$1.62$9.72
Gemini 3.1 ProFrontier multimodal$2.16$12.96
Gemini 3 FlashFast multimodal$0.540$3.24
Gemini 3.1 Flash-LiteCost-efficient steps$0.270$1.62

GPT-5.6 tiers, Kimi K2.6, and DeepSeek V4 Pro are served from Edgaze's managed Azure AI Foundry capacity on hosted runs. Claude models are served directly from Anthropic, Gemini models directly from Google.

Image Models#

ModelBest forPer image (standard)
Nano Banana 2RecommendedDefault image generation$0.072
Gemini 3 Pro ImagePremium image quality$0.145
Nano Banana 2 LiteHigh-volume, low cost$0.037
OpenAI Image 1 miniCheap, quick generations$0.022
ChatGPT Image 2Premium OpenAI image$0.086
OpenAI Image 1.5Balanced quality + cost$0.065

OpenAI image rates vary with size and quality settings; the estimator prices the configured variant.

Embedding Models#

ModelBest forPer 1M input tokens
text-embedding-3-smallRecommendedSemantic search + RAG$0.022
text-embedding-3-largePremium RAG retrieval$0.140

Bring Your Own Key (BYOK)#

BYOK runs call the provider directly on your API key. OpenAI-brand, Anthropic, and Google models support this, and the token cost lands on your provider account instead of the rates above (a flat orchestration fee applies; see BYOK). Kimi K2.6 and DeepSeek V4 Pro run exclusively on Edgaze's managed capacity and do not support BYOK.

Automatic Model Migration#

Providers retire models on their own schedules, and the Edgaze catalog evolves with them. You never need to hand-edit old workflows:

  • Selectable models appear in the builder dropdown and are fully supported.
  • Deprecated models disappear from the dropdown but keep running and billing at their listed rate. Existing workflows are unaffected.
  • Retired models are automatically routed to their designated successor within the same provider family at run time (for example, gpt-4o-mini → GPT-5.6 Luna, claude-sonnet-4-6 → Claude Sonnet 5, dall-e-3 → OpenAI Image 1.5). Cost estimates, holds, and billing all use the successor's rate, so what you're quoted always matches what actually runs.

Stored workflows are never bulk-edited: the remap happens when a run executes, and the builder writes the current id the next time you save a workflow.

Reliability#

Every chat and image model carries an ordered cross-provider failover chain. If a provider returns a transient error (rate limit, timeout, outage) mid-run, the engine can substitute the next model in the chain. The run's cost hold caps execution, and failover does not raise the displayed price quoted before the run.

Questions#

Contact [email protected] for pricing questions, or see Pricing & Limits for account-level limits.

Was this useful?

Your response helps us improve the documentation.

Related policies

Back to Legal & Trust Center
Infrastructure Cost Estimation

How the hosted compute estimate affects the price buyers see, without reducing creator margin earnings.

Workflow Run Policy

Per-run hosted workflow execution, BYOK behavior, funding sources, and abnormal usage controls.

Pricing Limits

Minimum and maximum per-run workflow margin on Edgaze.

Open Workflow BuilderOpen Workflow StudioBrowse workflow templatesVisit the marketplaceLearn about creators
On this page
OverviewChat ModelsImage ModelsEmbedding ModelsBring Your Own Key (BYOK)Automatic Model MigrationReliabilityQuestions
© 2026 Edge Platforms, Inc. All rights reserved.