Support
♧

Notifications are coming soon

◐
Console Support Sign in

One gateway for your AI models

Enter the API URL, your customer key and a model ID in compatible software to access verified models.

Limited trial API URLhttps://guko.si/api/v1

Explore leading model families

OpenAIOpenAIClaudeClaude
GPT-5.6GPT-6GPT-6.1Claude OpusClaude SonnetClaude HaikuClaude FableGPT Image
MODEL COLLECTION

Find the model for your next idea

Compare GPT, Claude and image models by use case and reference cost.

View available access ↗

20 modelsUpstream reference prices · USD

gpt-5.6-sol

OpenAI
gpt-5.6-sol
Input $3.75 / 1MOutput $22.50 / 1M

A GPT catalog version for conversation, coding assistance and writing. Evaluate each version and suffix separately.

Chat & codingText tested
Reference pricing details

openai-award · 2026-10-04 snapshot. Per million tokens at the base context tier, without cache or risk adjustments.

Check access & platform price

gpt-6.1-sol

OpenAI
gpt-6.1-sol
Input $1.875 / 1MOutput $11.25 / 1M

A GPT catalog version for conversation, coding assistance and writing. Evaluate each version and suffix separately.

Chat & codingUpstream catalog
Reference pricing details

openai-award · 2026-10-04 snapshot. Per million tokens at the base context tier, without cache or risk adjustments.

Check access & platform price

gpt-6-sol

OpenAI
gpt-6-sol
Input $1.875 / 1MOutput $11.25 / 1M

A GPT catalog version for conversation, coding assistance and writing. Evaluate each version and suffix separately.

Chat & codingUpstream catalog
Reference pricing details

openai-award · 2026-10-04 snapshot. Per million tokens at the base context tier, without cache or risk adjustments.

Check access & platform price

gpt-6-astra

OpenAI
gpt-6-astra
Input $7.5 / 1MOutput $37.50 / 1M

A GPT catalog version for conversation, coding assistance and writing. Evaluate each version and suffix separately.

Chat & codingUpstream catalog
Reference pricing details

openai-award · 2026-10-04 snapshot. Per million tokens at the base context tier, without cache or risk adjustments.

Check access & platform price

gpt-5.6-terra

OpenAI
gpt-5.6-terra
Input $1.875 / 1MOutput $11.25 / 1M

A GPT catalog version for conversation, coding assistance and writing. Evaluate each version and suffix separately.

Chat & codingUpstream catalog
Reference pricing details

openai-award · 2026-10-04 snapshot. Per million tokens at the base context tier, without cache or risk adjustments.

Check access & platform price

gpt-5.5

OpenAI
gpt-5.5
Input $3.75 / 1MOutput $22.50 / 1M

A GPT catalog version for conversation, coding assistance and writing. Evaluate each version and suffix separately.

Chat & codingText tested
Reference pricing details

openai-award · 2026-10-04 snapshot. Per million tokens at the base context tier, without cache or risk adjustments.

Check access & platform price

claude-opus-5

Claude
claude-opus-5
Input $30 / 1MOutput $150.00 / 1M

A Claude catalog version for document analysis, problem solving and writing. Available access is shown in the console.

Analysis & writingUpstream catalog
Reference pricing details

claude-award · 2026-10-04 snapshot. Per million tokens at the base context tier, without cache or risk adjustments.

Check access & platform price

claude-opus-5-5

Claude
claude-opus-5-5
Input $24 / 1MOutput $120.00 / 1M

A Claude catalog version for document analysis, problem solving and writing. Available access is shown in the console.

Analysis & writingUpstream catalog
Reference pricing details

claude-award · 2026-10-04 snapshot. Per million tokens at the base context tier, without cache or risk adjustments.

Check access & platform price

claude-opus-4-8

Claude
claude-opus-4-8
Input $30 / 1MOutput $150.00 / 1M

A Claude catalog version for document analysis, problem solving and writing. Available access is shown in the console.

Analysis & writingUpstream catalog
Reference pricing details

claude-award · 2026-10-04 snapshot. Per million tokens at the base context tier, without cache or risk adjustments.

Check access & platform price

claude-opus-4-7

Claude
claude-opus-4-7
Input $30 / 1MOutput $150.00 / 1M

A Claude catalog version for document analysis, problem solving and writing. Available access is shown in the console.

Analysis & writingUpstream catalog
Reference pricing details

claude-award · 2026-10-04 snapshot. Per million tokens at the base context tier, without cache or risk adjustments.

Check access & platform price

claude-opus-4-6

Claude
claude-opus-4-6
Input $30 / 1MOutput $150.00 / 1M

A Claude catalog version for document analysis, problem solving and writing. Available access is shown in the console.

Analysis & writingText tested
Reference pricing details

claude-award · 2026-10-04 snapshot. Per million tokens at the base context tier, without cache or risk adjustments.

Check access & platform price

claude-sonnet-5-5

Claude
claude-sonnet-5-5
Input $18 / 1MOutput $90.00 / 1M

A Sonnet-series option for coding assistance, document workflows and conversation. Select the full model ID.

Code & documentsUpstream catalog
Reference pricing details

claude-award · 2026-10-04 snapshot. Per million tokens at the base context tier, without cache or risk adjustments.

Check access & platform price

claude-sonnet-5

Claude
claude-sonnet-5
Input $18 / 1MOutput $90.00 / 1M

A Sonnet-series option for coding assistance, document workflows and conversation. Select the full model ID.

Code & documentsUpstream catalog
Reference pricing details

claude-award · 2026-10-04 snapshot. Per million tokens at the base context tier, without cache or risk adjustments.

Check access & platform price

claude-sonnet-4-6

Claude
claude-sonnet-4-6
Input $18 / 1MOutput $90.00 / 1M

A Sonnet-series option for coding assistance, document workflows and conversation. Select the full model ID.

Code & documentsUpstream catalog
Reference pricing details

claude-award · 2026-10-04 snapshot. Per million tokens at the base context tier, without cache or risk adjustments.

Check access & platform price

claude-haiku-4-5

Claude
claude-haiku-4-5
Input $12 / 1MOutput $60.00 / 1M

A Haiku-series option for everyday questions, classification and concise replies. Evaluate it on your own workload.

Everyday tasksUpstream catalog
Reference pricing details

claude-award · 2026-10-04 snapshot. Per million tokens at the base context tier, without cache or risk adjustments.

Check access & platform price

claude-fable-5

Claude
claude-fable-5
Input $60 / 1MOutput $120.00 / 1M

A Claude catalog version for document analysis, problem solving and writing. Available access is shown in the console.

Analysis & writingUpstream catalog
Reference pricing details

claude-award · 2026-10-04 snapshot. Per million tokens at the base context tier, without cache or risk adjustments.

Check access & platform price

claude-fable-5-1

Claude
claude-fable-5-1
Input $60 / 1MOutput $120.00 / 1M

A Claude catalog version for document analysis, problem solving and writing. Available access is shown in the console.

Analysis & writingUpstream catalog
Reference pricing details

claude-award · 2026-10-04 snapshot. Per million tokens at the base context tier, without cache or risk adjustments.

Check access & platform price

gpt-image-2

OpenAI
gpt-image-2
Reference $4.50 / request

Image generation catalog entry for illustration and visual concepts. Image access is not yet available on this platform.

Image generationUpstream catalog
Reference pricing details

openai-award · 2026-10-04 snapshot. Per-request reference only; image access is pending.

Check access & platform price

gpt-image-2.5-sunburst

OpenAI
gpt-image-2.5-sunburst
Reference $5.40 / request

Image generation catalog entry for illustration and visual concepts. Image access is not yet available on this platform.

Image generationUpstream catalog
Reference pricing details

openai-award · 2026-10-04 snapshot. Per-request reference only; image access is pending.

Check access & platform price

gpt-image-2.5-flare

OpenAI
gpt-image-2.5-flare
Reference $5.25 / request

Image generation catalog entry for illustration and visual concepts. Image access is not yet available on this platform.

Image generationUpstream catalog
Reference pricing details

openai-award · 2026-10-04 snapshot. Per-request reference only; image access is pending.

Check access & platform price

Reference prices are from the upstream catalog, not platform retail quotes. Unified access is pending; available models, permissions and platform prices are shown in the console. Catalog names do not confirm model provenance or that every model has been tested.

CORE SERVICES

From Model Inference to GPU Delivery—
Managed in One Platform.

Manage models, API keys, compute capacity, team access, and spend through one clear workflow.

API01

Unified Model API Access

Call multiple models through a consistent interface with centralized logs, quotas, and routing policies.

Multi-model access · Logs · Quotas
For AI products and SaaS teamsLearn more →
GPU02

On-Demand GPU Computing

Provision capacity for inference, training, deployment, and elastic scaling as requirements change.

Inference · Training · Deployment · Scaling
For ML and engineering teamsLearn more →
KEY03

API Key and Team Management

Control key permissions, member roles, project limits, and usage auditing from one place.

Permissions · Roles · Usage audits
For collaborative engineering teamsLearn more →
ENT04

Enterprise Infrastructure Solutions

Dedicated access, private or hybrid deployment, consolidated billing, and expert technical support.

Dedicated access · Security · Support
For enterprises and platform operatorsLearn more →
WHY TOKENFORGE

Lower Integration Overhead. Higher Delivery Velocity.

Eliminate duplicate platform work and keep your team focused on product quality and growth.

01

One API, Multiple Models

Manage access to multiple models with a single API key and integration pattern.

02

Intelligent Routing and Failover

Choose routes based on model health, latency, capability, and cost.

03

Production-Ready Concurrency

Scale request throughput with clearly defined capacity and integration support.

04

Transparent Usage Analytics

Track requests, token consumption, spend, and call logs in one view.

05

Elastic GPU Capacity

Configure GPUs for training, inference, deployment, and model evaluation.

06

Flexible Billing

Manage team balances, quotas, and consolidated enterprise billing.

07

Enterprise Technical Support

Access dedicated integration, private cloud, hybrid cloud, and deployment assistance.

08

Hands-On Regional Support

Work directly with a technical team that understands local and cross-border delivery requirements.

REGIONAL DEPLOYMENT

Regional Capacity. Flexible Routing.

Connect to multi-region resources and elastic scheduling options designed to improve reliability for AI applications across markets.

Automatic routingMulti-region resilienceRegional deployment
Hong KongAvailable · Auto-routed
SingaporeAvailable · Auto-routed
TokyoAvailable · Auto-routed
FrankfurtAvailable · Auto-routed
VirginiaAvailable · Auto-routed
DubaiAvailable · Auto-routed
Available regionsNETWORK VIEW
DEVELOPER FIRST

Integrate in Minutes

Use a familiar, OpenAI-compatible request format to reduce migration effort.

01Create an API KeyAssign project-specific access and quotas
02Select a ModelRoute by task, latency, and cost
03Send a RequestUse one format across apps and workflows
QUICK START

Keep the Workflow You Know

This sample illustrates the integration pattern. Production endpoints, available models, and credentials are confirmed during service onboarding.

INTEGRATION EXAMPLE · cURL
curl https://guko.si/api/v1/chat/completions \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "gpt-5.6-sol",
    "messages": [{"role": "user", "content": "Hello TokenForge"}]
  }'
PLATFORM CAPABILITIES

Platform Capabilities at a Glance

Review key capabilities across APIs, GPU scheduling, accounts, billing, and technical support. Production monitoring will be connected as services go live.

Platform Status OverviewSTATUS OVERVIEW
Model API ServicesOperational
GPU SchedulingOperational
Account ServicesOperational
Usage and BillingOperational
Technical SupportOperational
START BUILDING

Build Your AI Model and Compute Infrastructure

Share your expected model volume, GPU requirements, concurrency, and billing preferences. We will help define a practical path to production.