Fast & Efficient 1.05M Context

GPT-5.6 Luna

GPT-5.6 Luna is OpenAI's fast, cost-efficient model in the GPT-5.6 series. It handles text, image, and file inputs with a 1.05M token context — perfect for high-volume, latency-sensitive workloads on Hoodgen.

Provider: OpenAIStatus: Live

Technical Specifications

Live data fetched from the OpenRouter model registry — always up to date.

Context Length
Loading...
Entire repositories in one session
Max Completion
Loading...
Tokens per single response
Prompt Pricing
Loading...
Per million input tokens
Completion Pricing
Loading...
Per million output tokens
Modality
Loading...
Text + Image + Video input
Model ID
Loading...
OpenRouter identifier

What is GPT-5.6 Luna?

GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive applications while maintaining strong quality across text, image, and file inputs. With a 1.05M token context window, Luna can process large documents, codebases, and complex multi-modal inputs in a single session.

Features & Capabilities

Fast Responses

Optimized for low latency — ideal for real-time applications.

Cost Efficient

One of the most affordable models in the GPT-5.6 series.

Text + Image + File

Understands text, images, and file attachments natively.

1.05M Context

Process huge documents and codebases in one session.

Coding Capable

Strong code generation, review, and debugging.

Production Ready

Built for reliable, high-volume production workloads.

Who Is GPT-5.6 Luna For?

High-Volume Apps

Handle large numbers of requests with fast, consistent responses.

Document Processing

Analyze and extract from large documents and file attachments.

Coding Assistance

Generate, review, and debug code efficiently.

Multimodal Tasks

Work with text and images together for richer outputs.

GPT-5.6 Luna Pricing

₹0
starting from free credits
  • Unlimited chats on free plan credits
  • No credit card required
  • Included in Hoodgen AI free tier (100 credits/mo)
  • Premium models also available on paid plans
See Plans →

What Makes GPT-5.6 Luna Different?

DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model with 284B total parameters but only 13B active per token. This architecture is what makes it special — it delivers near-frontier quality at a fraction of the compute cost, which shows up as very low pricing on OpenRouter.

Three features define DeepSeek V4 Flash 0731's pitch. The first is the 1.3M token context window — enough to hold entire repositories and long task histories in one session. The second is the sparse MoE design that keeps both cost and latency low. The third is DeepSeek's open-weight philosophy, which has made its models widely adopted across the developer community.

DeepSeek V4 Flash 0731 also supports tool calling, structured outputs, and other agent-friendly features, making it a practical choice for production agentic workloads and automated pipelines.

GPT-5.6 Luna Performance

GPT-5.6 Luna is positioned as the speed-and-value tier of OpenAI's GPT-5.6 family. It delivers strong general-purpose performance — coding, analysis, multimodal understanding — at lower latency and cost than the Pro-tier models. For high-volume workloads, it is one of the most practical choices on the market.

Model
DeepSWE Score
Context
Price
GPT-5.6 Luna
Sparse MoE · 13B active
1.3M tokens
~$0.06/M input
DeepSeek V4 Flash
Fast MoE
GLM-5.3
62%
GPT-5.6-sol
52%

What Can You Do With GPT-5.6 Luna?

GPT-5.6 Luna excels at high-volume, latency-sensitive work. Build chatbots, process documents, generate code, and handle multimodal inputs — all with fast, cost-efficient responses. Its 1.05M context window means even large inputs fit in a single session.

Best Use Cases

  • High-volume chatbot and customer support
  • Document analysis and extraction
  • Code generation and review
  • Multimodal understanding (text + images)
  • Latency-sensitive production apps

GPT-5.6 Luna Privacy & Data Policy

GPT-5.6 Luna is developed by OpenAI and hosted on OpenRouter. Prompts and completions are processed by OpenAI's API. As with any third-party model, avoid sharing sensitive personal data.

Because GPT-5.6 Luna is a third-party hosted model, avoid pasting passwords, personal data, or proprietary source code you would not share with any external provider. Hoodgen AI always keeps free alternative models available so you can switch without interruption.

GPT-5.6 Luna vs Other Models

Feature
GPT-5.6 Luna
GPT-5.6 Luna
Claude Haiku 4.5
Type
Fast MoE
General
General
Context
1.3M tokens
1M
1M
Best For
High-volume, coding, long context
Fast, cost-efficient chat
Balanced general chat
Price
~$0.06/M
$1.25/M
$1.25/M
Vision
Text

Frequently Asked Questions

Yes. GPT-5.6 Luna is included in your Hoodgen AI free plan credits and available on paid plans.

Ready for fast, efficient AI?

Chat with GPT-5.6 Luna free — sign in and start a conversation in seconds.

Start Chatting Free