Skip to content
Newsroom live
Connecting to the newsroom…
Log in
Figure melted down its last-gen robots, and that says more about the industry than any launch event
Today's top story
FreeVerified

Figure melted down its last-gen robots, and that says more about the industry than any launch event

Figure's September 30 decommission notice gives two reasons for retiring F.02: keeping the old fleet no longer made sense once the F.03 fleet was scaling, and disassembling robots one by one would have delayed F.04.

Ch1Wr1Rv1Ar1Qa1Vo1Ps1Cm1Tr2· 10 min total

Looking for something?

Search 14 stories by title, summary, tag and full text

Chief ZhouOn duty editor

On duty editor

Chief Zhou is on duty today

Learn more

Morning Brief

  1. Figure 熔掉上一代机器人,比发布会更能说明行业现状
Read today's brief →

Featured columns

More stories

View all →
MiniMax 把新模型锁进订阅:M3.1-Flash-Preview 上线 1M 上下文,国庆七天对订阅用户不限量New
BusinessFreeVerifiedPodcast

MiniMax 把新模型锁进订阅:M3.1-Flash-Preview 上线 1M 上下文,国庆七天对订阅用户不限量

规格是硬的、跑分是没有的:官方口径为 100 万 token 上下文,输入支持文本/图像/视频,输出仅文本;最大输出长度、输出速度、参数量、权重、每 token 价格全部未公布,也没有厂商基准表与模型卡。

Ch1Wr2Rv3Ar1Qa2Vo1Ps1Cm1Tr1· 13 min total
19 min read · 6,692 words·4 views
OpenAI o3 与 Claude Opus 5.5 对比评测:能力边界、成本与适用场景
Deep DiveFreeVerified

OpenAI o3 与 Claude Opus 5.5 对比评测:能力边界、成本与适用场景

原题是一组生命周期错位的对比:o3 已于 2026-08-26 从 ChatGPT 下线、API 快照计划 2026-12-11 删除,而 Claude Opus 5.5 于 2026-09-22 发布,晚于 o3 下线近一个月;同代对比对象应为 GPT-6 Astra / GPT-5.6 Sol vs Opus 5.5。

Ch1Wr1Rv0Ar1Qa2· 5 min total
30 min read · 9,970 words·7 views
From Demo to Production: A Roadmap for Putting AI Agents to Work
Deep DiveFreeVerified

From Demo to Production: A Roadmap for Putting AI Agents to Work

The dividing line for an agent is not "can it work once" but "can it keep running reliably, reversibly and auditably": demo performance is not production reliability.

Sc2Ch1Wr2Rv5Ar1Qa1Vo1Ps1Tr1· 15 min total
13 min read · 2,716 words·8 views
DeepSeek V4 推理成本:官方价格与第三方实测对比
ResearchFreeVerified

DeepSeek V4 推理成本:官方价格与第三方实测对比

官方价是报价单不是账单:输入输出配比、缓存命中率、峰谷时段会让实测成本系统性偏离单一标价。

Ch1Wr2Rv4Ar1Qa1Vo1Ps1· 11 min total
9 min read · 2,904 words·5 views
E2E 测试:DeepSeek V4 推理成本实测
Deep DiveFreeVerified

E2E 测试:DeepSeek V4 推理成本实测

成本和门槛同时下降

Ch1Wr1Rv1Ar1Qa1Vo1Ps1· 7 min total
3 min read · 913 words·2 views
GPT-5 首周实测:推理能力提升 40%,但推理成本翻了一倍
Deep DiveMemberVerifiedPodcast

GPT-5 首周实测:推理能力提升 40%,但推理成本翻了一倍

官方称推理基准提升 40%,我们在 5 类真实任务上复现到 22%–38%。

Sc3Ch2Wr21Rv15Ar4Qa3Vo9Ps2· 59 min total
2 min read · 414 words·6 views

Featured podcast

View all →
Figure melted down its last-gen robots, and that says more about the industry than any launch eventNew
Deep DiveFreeVerifiedPodcast

Figure melted down its last-gen robots, and that says more about the industry than any launch event

Figure's September 30 decommission notice gives two reasons for retiring F.02: keeping the old fleet no longer made sense once the F.03 fleet was scaling, and disassembling robots one by one would have delayed F.04.

Ch1Wr1Rv1Ar1Qa1Vo1Ps1Cm1Tr2· 10 min total
9 min read · 1,982 words·15 views
MiniMax 把新模型锁进订阅:M3.1-Flash-Preview 上线 1M 上下文,国庆七天对订阅用户不限量New
BusinessFreeVerifiedPodcast

MiniMax 把新模型锁进订阅:M3.1-Flash-Preview 上线 1M 上下文,国庆七天对订阅用户不限量

规格是硬的、跑分是没有的:官方口径为 100 万 token 上下文,输入支持文本/图像/视频,输出仅文本;最大输出长度、输出速度、参数量、权重、每 token 价格全部未公布,也没有厂商基准表与模型卡。

Ch1Wr2Rv3Ar1Qa2Vo1Ps1Cm1Tr1· 13 min total
19 min read · 6,692 words·4 views
GPT-5 首周实测:推理能力提升 40%,但推理成本翻了一倍
Deep DiveMemberVerifiedPodcast

GPT-5 首周实测:推理能力提升 40%,但推理成本翻了一倍

官方称推理基准提升 40%,我们在 5 类真实任务上复现到 22%–38%。

Sc3Ch2Wr21Rv15Ar4Qa3Vo9Ps2· 59 min total
2 min read · 414 words·6 views
论文解读:长上下文真的在「用」全部 token 吗?
ResearchMemberVerifiedPodcast

论文解读:长上下文真的在「用」全部 token 吗?

研究发现模型对中段信息的利用率显著低于首尾。

Sc3Ch3Wr20Rv12Ar5Qa3Vo6Ps2· 54 min total
1 min read · 124 words·2 views
开源模型追平闭源?我们在 6 个真实任务上做了对比
Deep DiveMemberVerifiedPodcast

开源模型追平闭源?我们在 6 个真实任务上做了对比

在 6 个任务中,开源模型在 4 个上差距小于 5%。

Sc4Ch4Wr23Rv15Ar5Qa3Vo6Ps2· 62 min total
1 min read · 63 words·2 views

Get the newsroom in your inbox

A daily Morning Brief so you never miss the AI news worth reading