G
OpenAIDeprecated

GPT-4 Turbo API

Previous-generation model. Use GPT-4.1 or GPT-4o for new projects.

💰 Save up to 70% vs official OpenAI pricing
TL;DR
Price: $10.00/1M input · $30.00/1M output
Context: 128K · max 4,096 output
Provider: OpenAI
Cost advantage: Cheaper than official API · No credit card

GPT-4 Turbo — cheaper than the official OpenAI API

Access GPT-4 Turbo through AI API Hub and pay less per token. Same OpenAI-compatible endpoint, lower cost.

$10.00/1M
input token price
INPUT / 1M tokens
$10.00
OUTPUT / 1M tokens
$30.00
CONTEXT WINDOW
128K

Technical Specifications

ProviderOpenAI
Model FamilyGPT-4 Turbo
Release Date2023-11
Context Window128K
Max Output Tokens4,096
Input Price$10.00 / 1M tokens
Output Price$30.00 / 1M tokens
Vision SupportYes ✓
Function CallingYes ✓
JSON ModeYes ✓
StreamingNo
Fine TuningAvailable
StatusDeprecated ⚠

Overview

GPT-4 Turbo is OpenAI's previous-generation gpt4 model, released in 2023-11. Previous-generation model. Use GPT-4.1 or GPT-4o for new projects.

For developers, the headline numbers are a 128K context window and up to 4,096 output tokens per response — enough headroom for 128k context and vision capable without chunking your input. Priced at $10.00/1M input and $30.00/1M output, it sits in the premium tier — best reserved for tasks where quality justifies the spend.

On the capability side, GPT-4 Turbo exposes 3 features: Vision, Function Calling, JSON Mode. Fine-tuning is on the table if you need to specialize behavior on your own data. Vision support means you can pass images alongside text, handy for document parsing or UI automation.

The practical appeal of routing GPT-4 Turbo through AI API Hub is simplicity: one OpenAI-compatible endpoint, USDT & USDC payments, no credit card, and you're calling the API in under 30 seconds — just swap your base URL.

What Makes GPT-4 Turbo Different

How GPT-4 Turbo is used

GPT-4 Turbo is used for agentic workflows combining visual understanding with action — document processing pipelines that extract data and call APIs, visual Q&A systems, and multimodal agents. Tool calling enables it to trigger external functions based on what it sees in images.

Pricing position within OpenAI

GPT-4 Turbo sits in the middle of OpenAI's pricing at $10.00/1M input — 230% above the lineup average ($0.10 cheapest, $20.00 most expensive). 11 siblings cost less, 1 cost more. This mid-tier positioning makes it a sensible default when you're unsure which variant to pick.

GPT-4 Turbo's role in the lineup

Within OpenAI's lineup, GPT-4 Turbo is a mid-tier option — balanced between cost and capability. The gpt4 family has 5 active variants, and GPT-4 Turbo occupies the upper end. This makes it a safe default for production workloads where you're not sure which tier to pick.

Real-world use cases

Real-world deployments: document processing pipelines (read invoice → extract fields → call accounting API), visual Q&A systems, and multimodal agents that act on what they see. GPT-4 Turbo handles the full see-decide-act loop in a single model call.

vs sibling models

What makes GPT-4 Turbo different from sibling models: compared to GPT-5.5 ($5.00/1M cheaper, 256K vs 128K context (larger)); GPT-5.4 ($7.50/1M cheaper, 256K vs 128K context (larger)); GPT-4.1 ($8.00/1M cheaper, 1M vs 128K context (larger)). Choose GPT-4 Turbo when vision input is needed.

API Examples

Python

from openai import OpenAI

client = OpenAI(
    api_key="YOUR_API_KEY",
    base_url="https://api.apiyihe.org/v1"
)

response = client.chat.completions.create(
    model="gpt-4-turbo",
    messages=[
        {"role": "user", "content": "Hello"}
    ]
)

print(response.choices[0].message.content)

JavaScript / Node.js

import OpenAI from "openai";

const client = new OpenAI({
  apiKey: process.env.API_KEY,
  baseURL: "https://api.apiyihe.org/v1"
});

const response = await client.chat.completions.create({
  model: "gpt-4-turbo",
  messages: [
    { role: "user", content: "Hello" }
  ]
});

console.log(response.choices[0].message.content);

cURL

curl https://api.apiyihe.org/v1/chat/completions \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -d '{
    "model": "gpt-4-turbo",
    "messages": [
      {"role": "user", "content": "Hello"}
    ]
  }'

Supported Features

Vision / Image Input✅ Supported
Audio / Voice Input❌ Not Available
Function Calling✅ Supported
JSON Mode✅ Supported
Streaming❌ Not Available
Fine-Tuning✅ Available
Multimodal❌ Not Available

Benchmark Scores

BenchmarkScore
MMLUNot Publicly Available
GPQANot Publicly Available
SWE-BenchNot Publicly Available
HumanEvalNot Publicly Available
GSM8KNot Publicly Available
MATHNot Publicly Available
MMMUNot Publicly Available
Scores are from official provider publications. Empty fields indicate benchmarks not yet publicly disclosed.

Pricing History

GPT-4 Turbo was released in 2023-11 by OpenAI and is currently deprecated — consider upgrading to a newer model.

Current Pricing: $10.00 per 1M input tokens · $30.00 per 1M output tokens. Pay-as-you-go with no minimum commitment.

Pricing Model: Token-based billing (pay per use). No subscription fees. No hidden costs. Fine-tuning incurs additional costs at training and inference time.

💡 OpenAI occasionally updates pricing. AI API Hub reflects current pricing in real-time. All prices in USD. Pay with USDT or USDC — no currency conversion fees.

Compare Alternatives

Frequently Asked Questions

What is GPT-4 Turbo?

GPT-4 Turbo is OpenAI's previous-generation gpt4 model. Previous-generation model. Use GPT-4.1 or GPT-4o for new projects. It offers a 128K context window and supports Vision, Function Calling, JSON Mode. You can access it through AI API Hub using USDT or USDC — no credit card required.

How much does GPT-4 Turbo cost?

GPT-4 Turbo is priced at $10.00 per 1M input tokens and $30.00 per 1M output tokens, billed pay-as-you-go with no minimum. Through AI API Hub you can start with as little as $5 and scale from there.

GPT-4 Turbo vs Claude Opus 4.8?

They're built for different jobs. GPT-4 Turbo costs $10.00/1M input with a 128K window; Claude Opus 4.8 runs $5.00/1M input with 1M. Claude Opus 4.8 is the cheaper option, while GPT-4 Turbo trades cost for 128k context. See the full side-by-side at /compare/gpt-4-turbo-vs-claude-opus-4.8/.

GPT-4 Turbo context window?

GPT-4 Turbo has a 128K context window, capable of processing up to 128,000 tokens in a single request. Maximum output tokens: 4,096.

Does GPT-4 Turbo support function calling?

Yes, GPT-4 Turbo supports function/tool calling, allowing you to define functions that the model can invoke. This enables AI agents, API integrations, and structured data extraction.

Is GPT-4 Turbo multimodal?

Partially — GPT-4 Turbo supports vision (image input) but not native audio processing.

GPT-4 Turbo API rate limits?

GPT-4 Turbo rate limits: 10K RPM. Higher tier plans offer increased throughput. For high-volume production use, consider OpenAI's faster variant models.

How to access GPT-4 Turbo API?

Access GPT-4 Turbo through AI API Hub: (1) Register at api.apiyihe.org/register?aff=8JZC, (2) Deposit USDT/USDC, (3) Get your API key instantly, (4) Use the OpenAI-compatible endpoint https://api.apiyihe.org/v1 with model name "gpt-4-turbo". Start building in under 30 seconds.

Get GPT-4 Turbo API Access

Pay with USDT & USDC. Same model, up to 70% less.

계정 만들기
API 키 받기