24 Downloads Updated 4 days ago
ollama run treyleo16/gpt-5-6-luna
Updated 4 days ago
4 days ago
347016434c3d · 128kB ·
GPT-5.6 Luna is OpenAI’s high-speed, cost-optimized model tier within the GPT-5.6 family (alongside Sol and Terra). Serving as the lightweight nano-tier engine of the suite, Luna is engineered for high-throughput, latency-sensitive applications, high-volume classification, and streaming sub-agent orchestration.
Luna bridges the gap between ultra-low unit economics and frontier-class context processing. It enables developers to run continuous, large-scale agentic pipelines and full-document passes without incurring the cost overhead of flagship reasoning engines.
| Parameter | Specification |
|---|---|
| Model ID | gpt-5.6-luna / openai/gpt-5.6-luna |
| Context Window | 1,100,000 tokens (1.1M) |
| Max Output Tokens | 272,000 tokens |
| Input Modalities | Text, Code, Images (Vision) |
| Output Modalities | Text, Code, Structured JSON |
| Native Features | Function Calling, Prompt Caching, Reasoning Tokens, Computer Use, MCP Tools |
| Tier | Role & Capability Target |
|---|---|
| GPT-5.6 Sol | Flagship: Maximum reasoning, top-tier coding, defensive cybersecurity, and deep multi-step architecture. |
| GPT-5.6 Terra | Balanced Workhorse: Mid-tier option balancing intelligence and cost for standard enterprise production traffic. |
| GPT-5.6 Luna | Ultra-Fast Efficiency: High-throughput, low-latency engine designed for cost-sensitive, high-volume workloads. |