gpt-6-astra 出现特殊 API 响应,OpenAI 下一代模型 Astra 或为 GPT-6
作者转发 @synthwavedd 的发现:OpenAI Responses API 对 "gpt-6-astra" 返回 404,而无效 slug 返回 400,已知存在的 5.6 Cyber 也返回 404,此前 Bedrock API 曾以类似特征提前一天暴露 Fable 5.1。
X 热议
分类页是滚动库存,每页 24 条。首页只放当天精选。
作者转发 @synthwavedd 的发现:OpenAI Responses API 对 "gpt-6-astra" 返回 404,而无效 slug 返回 400,已知存在的 5.6 Cyber 也返回 404,此前 Bedrock API 曾以类似特征提前一天暴露 Fable 5.1。
Google DeepMind 发布两款新 Gemini 模型:Gemini 3.8 Flash 相比 3.7 Flash 在软件工程、智能体任务和多步推理上有显著提升;3.8 Flash Cyber 面向网络安全,提供漏洞检测与自动修补能力。
Google 发布 Gemini 3.8 Flash Cyber,定位为其最强网络安全模型,具备漏洞发现与大规模修补能力,速度和定价对标 Flash 档。模型在 CyberGym 基准达到 86.2%,CWE-Bench 修补任务 47.2%,内部基准上跨 20 种编程语言的漏洞发现成功率超过 70%。
Gemini 3.8 Flash 在 DeepSWE 得分 73%,定价表现引作者热议
阑夕实测发现 GPT-Image-2 偶尔会有思维链残余:一次生成图片触发风控被拒绝,但已生成的草图在尚未按其要求切割成 5 张图片之前,仍保留在步骤里。
Google 发布 Gemini 3.8 Flash,为 6 周内第 3 次 Flash 更新。官方称相较 3.7 Flash 在软件工程、智能体任务和多步推理上有显著提升,在 DeepSWE v1.1 上以更低成本自主端到端解决复杂工程问题,表现超越多数更大的前沿模型。图中基准显示其 Terminal-bench 2.1 得 89.4%、HLE-Verified 得 54.9%,输入价格 $0.75/1M tokens、输出 $3.…
Google 发布 Gemini 3.8 Flash,这是 6 周内第 3 次更新的 Flash 模型,官方称智能体与编码能力再次提升。
OpenRouter 宣布 Google DeepMind 的 Gemini 3.8 Flash 已上线,入口为 https://openrouter.ai/google/gemini-3.8-flash。官方称其以 3.7 Flash 的引入价在所列各项基准(编码、金融、法律、视频、科学)上超越前代;配图显示输入/输出价为 $0.75/$3.75 每百万 token(常规价 $1.50/$7.50),引入价有效期至 2026 年 1…
Gemini 3.8 Flash 在 DeepSWE V1.1 基准上得分 73.7%。配套图表显示该成绩在多款模型的分数与每任务平均成本对比中位于高效区,图表来源标注为 Datacurve AI(deepswe.datacurve.ai)。
通义千问发布 Qwen3.8-Max-0902,在 Code Arena: WebDev 以 1,691 分首次亮相即排名总榜第一,并以混合价 $5/MToken 成为 Pareto 前沿上得分最高的模型,现已可在 QwenCloud 试用。
商汤发起 U1.5 Lite AI 生成挑战赛,奖品含礼品卡与 U1 Pro Beta 权限,推自家生成模型。
官方账号:LongCat-2.0 进 Cline,装好选模型就能换。
RT @danshipper: BREAKING: Anthropic just dropped Fable 5.1—and CLAUDE IS SO BACK. We’ve spent the last week testing it at @every across c…
RT @claudeai: We’re introducing Claude Fable 5.1 and Claude Mythos 5.1. They're the world’s most advanced models for coding and knowledge…
Artificial Analysis 评测 Claude Fable 5.1,其在 max effort 下得 66 分登顶 Artificial Analysis Intelligence Index。
Rohan Paul 梳理了 Fable 5.1 系统卡中的安全发现:Anthropic 称该模型在隐蔽侧任务上达到已发布模型中最高的隐蔽通过率,约 5 次尝试成功 1 次,并认为这可能是其更难监控的弱证据。
With Fable 5.1 out today, we've also reset 5-hour and weekly limits for all users.
Fable 5.1 is here! Beats Fable 5 across the board, and is SOTA on coding and knowledge work. Also scores much better on cost per task. More broadly, we’ve heard your feedback aroun
RT @claudeai: We’re introducing Claude Fable 5.1 and Claude Mythos 5.1. They're the world’s most advanced models for coding and knowledge…
We’re introducing Claude Fable 5.1 and Claude Mythos 5.1. They're the world’s most advanced models for coding and knowledge work. https://t.co/8P9PSrWPi3
ZHO 转发 @takenoko_vr:本以为 AI 会变成插画/作曲/建模的入口,结果多数人停在生成、对成品没有违和感。
混元宣布 Hy4 preview:约 770B 参数、1M 上下文,并展示动画/角色生成效果,面向创作者试用。
面壁智能转发:日本创作者用 VoxCPM 语音生成,不露脸不出声把 YouTube 订阅做到约 1800,并觉得联盟营销赛道对手少。