Comparison

AI API Cost Calculator: Integrated Comparison of OpenAI·Claude·Gemini (as of June 2024)

ANSWERCompares real‑time pricing, features, and limits of OpenAI·Claude·Gemini APIs and gives cost‑calculation examples. Information reflects the latest official data as of June 2024.

Comparison of Pricing Policies by Major AI API Providers

OpenAI API (GPT‑4o, GPT‑4 Turbo, GPT‑3.5 Turbo)

  • Input token price: GPT‑4o (as of May 2024) 0.005 USD per 1K tokens, GPT‑4 Turbo 0.01 USD per 1K tokens, GPT‑3.5 Turbo 0.0005 USD per 1K tokens
  • Output token price: GPT‑4o 0.015 USD per 1K tokens, GPT‑4 Turbo 0.03 USD per 1K tokens, GPT‑3.5 Turbo 0.0015 USD per 1K tokens
  • Free trial: $5 worth of free credits (usable for 90 days)
  • Limits: Requests per second (10–30 RPS), daily request limit (10,000–100,000 requests)
  • Recommended for: Complex natural‑language processing, code generation, multilingual support

Claude API (Claude 3 Opus, Sonnet, Haiku)

  • Input token price: Claude 3 Opus 0.015 USD per 1K tokens, Sonnet 0.003 USD per 1K tokens, Haiku 0.00025 USD per 1K tokens
  • Output token price: Claude 3 Opus 0.075 USD per 1K tokens, Sonnet 0.015 USD per 1K tokens, Haiku 0.00125 USD per 1K tokens
  • Free trial: $5 worth of free credits (usable for 90 days)
  • Limits: Requests per second (5–20 RPS), daily request limit (5,000–50,000 requests)
  • Recommended for: Real‑time conversation, document summarization, code review

Google Gemini API (Gemini 1.5 Pro, Flash)

  • Input token price: Gemini 1.5 Pro 0.0075 USD per 1K tokens, Flash 0.000125 USD per 1K tokens
  • Output token price: Gemini 1.5 Pro 0.03 USD per 1K tokens, Flash 0.000375 USD per 1K tokens
  • Free trial: $3 worth of free credits (usable for 90 days)
  • Limits: Requests per second (10–50 RPS), daily request limit (10,000–100,000 requests)
  • Recommended for: Multimodal (text + image) processing, large‑scale document analysis

Cost‑Calculation Examples and Savings Tips

Practical Cost‑Calculation Method

  1. Count tokens: Add the tokens in the request text and the response text. (Example: 1,000 words ≈ 1,500 tokens)
  2. Apply prices: Multiply the number of input and output tokens by the respective model prices.
    • Example 1) Using GPT‑4o for a 500‑word request (input) and a 1,000‑word response (output)
      • Input: 750 tokens × 0.005 USD = 0.00375 USD
      • Output: 1,500 tokens × 0.015 USD = 0.0225 USD
      • Total: 0.02625 USD (≈ 35 KRW)
    • Example 2) Using Gemini 1.5 Flash to analyze a 10,000‑word document (input only)
      • Input: 15,000 tokens × 0.000125 USD = 0.001875 USD
      • Output: 0 (only analysis result requested)
      • Total: 0.001875 USD (≈ 2.5 KRW)
  3. Estimate monthly cost: Multiply daily request count by 30 days and by the average cost per request.

Savings Tips

  • Model selection: Use GPT‑4o for complex tasks, GPT‑3.5 Turbo or Gemini Flash for simple tasks
  • Caching: Store identical responses locally to avoid repeated calls
  • Token optimization: Simplify prompts and minimize unnecessary output
  • Leverage free tier: Test within the free‑credit period before switching to paid usage

Feature Comparison

FeatureOpenAI (GPT‑4o)Claude (Sonnet)Google (Gemini 1.5 Pro)
Code generationExcellent (Python, JS, etc.)Good (Python, Go)Fair (Python, C++)
Document summarizationGood (long‑form)Best (structured)Best (multimodal)
Real‑time conversationFairBest (fast response)Fair
Image processingLimitedLimitedBest (Gemini Vision)
Multilingual support80+ languages100+ languages100+ languages

Limits and Reliability

  • Request rate (RPS): Claude Haiku 5 < OpenAI GPT‑3.5 10 < Gemini Flash 50
  • Daily request limit: Claude Opus 5,000 < OpenAI GPT‑4 10,000 < Gemini Pro 100,000
  • Latency: Claude 3 Opus 1–2 s < GPT‑4o 1–3 s < Gemini 1.5 Pro 2–5 s

Model Choice by Target Audience

  1. Developers / code generation: OpenAI GPT‑4o (high code quality) or Claude Sonnet (fast feedback)
  2. Enterprise document analysis: Google Gemini 1.5 Pro (multimodal, large‑scale processing)
  3. Customer‑service chatbots: Claude Haiku (low cost, fast response)
  4. Multilingual services: Claude 3 (100+ languages) or Gemini (integrated with Google ecosystem)
  5. Budget‑conscious projects: Gemini Flash (lowest input token price) or GPT‑3.5 Turbo (cheap output tokens)

Sources checked

openai.comOpenAI | Research & Deployment (2026-08-12)

gemini.google.com‎Google Gemini (2026-08-12)

chatgpt.comChatGPT: Chat, Work, Create & Code with AI (2026-08-12)

gemini.google.comGoogle Gemini (2026-08-12)

chatgpt.comChatGPT 플랜 | 무료 (2026-08-12)

m.blog.naver.com인공지능 (AI)이란? 기본 개념부터 미래 전망까지 완벽 정리 ... (2026-08-12)

Open provider document
Next guideComparison of Monthly Plans for Meeting Minutes AI Services