AIHubMix · AIHubMix platform · LLM

gpt-5.6-luna

Release announcedReleased on Jul 9, 2026

Notice

Release announced

Released on Jul 9, 2026

AIHubMix announced this model on 2026-07-09.

Hosting route
AIHubMix platform
Affected scope
Not stated in the notice
Announced
Jul 9, 2026
First seen by ModelClock
Oct 5, 2026

What the provider published

docs.aihubmix.com ↗

​ GPT-5.6 系列与提示词缓存更新

​ 新增模型

  • 新增 gpt-5.6-sol、gpt-5.6-terra、gpt-5.6-luna 三款 GPT-5.6 系列模型(OpenAI 2026-07-09 正式发布)。三档均为 1,050,000 上下文窗口、128K 最大输出、知识截止 2026-02-16,支持文本与图像输入,可通过 Chat Completions、Responses 与 Claude 兼容的 Messages 接口调用。Sol 为旗舰档,面向复杂专业工作,官方称其为当前最佳编码模型;Terra 性能与 GPT-5.5 相当且价格减半;Luna 面向成本敏感场景。

​ GPT 提示词缓存文档上线

  • 新增 GPT 提示词缓存 文档:GPT-5.6 系列起缓存写入按 1.25 倍输入价计费、缓存读取按 0.1 倍计费、缓存至少保留 30 分钟;覆盖 prompt_cache_key 与显式缓存断点参数说明、计费逻辑、接口示例与命中排查。 提示词缓存实践 与 Claude 提示词缓存 的 OpenAI 缓存口径同步更新。

​ Claude Fable/Mythos 思考模式兼容

  • Claude Fable 与 Mythos 系列在请求中使用 reasoning_effort 时,现在会按 adaptive thinking 处理,减少这类模型因思考参数不匹配导致的模型推理厂商 400。

​ Claude 与 Gemini 的 stop 参数生效更一致

  • OpenAI 兼容请求里的 stop 现在会映射到 Claude 与 Gemini 原生请求。Claude 不接受的纯空白 stop 会被过滤,OpenAI/Gemini 的数量限制也按各自规则处理,跨协议生成截断更稳定。

​ gpt-chat-latest token 上限参数兼容

  • gpt-chat-latest 以及后续 GPT/ChatGPT latest 别名在需要时会保留 max_completion_tokens,减少因误发旧字段 max_tokens 导致的 400。

​ Key 额度耗尽提示更清晰

  • 当 Key 级别使用限额耗尽时,API 现在返回更明确的引导文案,提示去调整并激活 Key 限额,而不是仅返回底层 quota 报错。

​ Tool calls 响应结构对齐 OpenAI

  • 非流式 chat 响应在只有 tool_calls、没有文本内容时,现在会显式返回 content: null,提升 OpenAI 风格 SDK 与响应解析器的兼容性。

​ Release 2026.07.09

Read from the provider's text; the quote is the provider's exact lines.

Also published on docs.aihubmix.com

Revision history

Loading…
Raw history JSON

Same model elsewhere

Each provider and hosting route sets its own dates.