社区讨论 · 赛道

EastVoice · AI · 2026-10-05 · Issue #88

EastVoiceEastVoice1 小时前2026/10/04 11 浏览

Editor's Note: China's AI week was about closing gaps at speed. StartLux open-sourced a decision model it says beats Jev on 31 of 38 benchmarks; Kimi K3 became the first Chinese model billed through OpenAI's enterprise system; and 22 researchers, Hinton and Bengio among them, warned that AI building AI is accelerating toward an intelligence explosion.

Commentary: The Chinese story this week is not one breakthrough but a steady drumbeat. Models ship, prices fall, and the country's open weights keep handing rivals free benchmarks.


I. AI Foundation Models

1. Anthropic Is Reportedly Courting the Vatican, and It Isn't Alone

  • Summary: Several foreign outlets report that Anthropic, the lab behind Claude, has been quietly lobbying the Vatican, one of the stranger crossovers of 2026. Top labs are now chasing influence well beyond the usual policy rooms.
  • Source: TMTPost · 2026-10-04
  • Editor's Take: Once labs run out of benchmark bragging rights, they start competing on legitimacy. Religion is an old institution with a long memory and a global audience.

2. After Jev, Chinese Teams Dig Into AI's "Intuition Layer"

  • Summary: The pitch is simple: let the big model think, let a small model judge. According to GeekPark, Chinese teams are racing to build that fast decision layer, weeks after a former OpenAI researcher's work set off the debate.
  • Source: GeekPark · 2026-10-04
  • Editor's Take: Splitting thought from judgment is how you make an agent cheap enough to run always-on. The race here is latency per decision, not parameter count.

3. When AI Starts Building AI

  • Summary: TMTPost traces how AI-assisted research keeps compressing the loop between one model design and the next, framing it as another handoff in a chain that spans generations rather than a single leap.
  • Source: TMTPost · 2026-10-04
  • Editor's Take: Self-improvement is the story everyone discusses and few can measure. The measurable part right now is cycle time, and it keeps shrinking.

4. StartLux Open-Sources a Decision Model, Claiming 31 of 38 Wins Over Jev

  • Summary: Shanghai-based StartLux released StartLux-Decision on September 30 in five sizes — 0.8B, 2B, 4B, 9B and 27B — and says it beats TypeSafe AI's Jev 1.13 on 31 of 38 benchmarks.
  • Source: TMTPost · 2026-10-04
  • Editor's Take: Shipping five sizes at once is a distribution play, not a research result. Owning every tier from phone to cluster is how you force rivals to chase your weights.

5. An OpenAI Veteran's Exit Letter Lays Bare a Culture Problem

  • Summary: A departing OpenAI employee wrote in The Atlantic about why the company's habit of framing every release as a nuclear-grade event finally wore thin. HuXiu picked up the essay and its implications for the lab's culture.
  • Source: HuXiu · 2026-10-04
  • Editor's Take: Hype is a recruiting tool until it becomes a retention problem. The people best at measuring a claim tend to be the first to leave when it stops matching reality.

6. AI Office Work Has Only Just Reached 1972

  • Summary: A long analysis argues today's AI office tools sit at roughly the spreadsheet-and-word-processor stage of maturity — useful, fragmented, and far from replacing a workflow end to end.
  • Source: HuXiu · 2026-10-04
  • Editor's Take: "AI office" is being sold as finished software while it remains a pile of features. Whoever turns that pile into one boring, reliable workflow wins.

7. Musk's Terafab Still Leans on Chinese Manufacturing

  • Summary: On October 3 Elon Musk confirmed a hybrid plan for Terafab: Intel runs the advanced 14A front-end process, while TSMC handles back-end operations, yield and packaging. Chinese media read it as Musk still trusting Chinese manufacturing capacity.
  • Source: QbitAI · 2026-10-04
  • Editor's Take: Splitting a fab across two foundries is a supply-chain workaround. Whoever controls back-end yield quietly controls the schedule.

8. Why China Isn't Getting Its Own "Muse"

  • Summary: A TMTPost essay argues the US browser-agent race does not transfer to China. Kimi, ByteDance and Tencent each keep users inside their own walls, so there is nothing for a single agent to unlock the way a browser unlocks the open web.
  • Source: TMTPost · 2026-10-04
  • Editor's Take: Agent strategy follows market structure, not model quality. Copy the product into a walled garden and you copy the shelved version of it.

9. Kimi K3 Becomes the First Chinese Model in OpenAI's Enterprise Billing

  • Summary: US inference provider Baseten said on September 30 that enterprises can run Moonshot AI's Kimi K3 inside OpenAI's Codex, billed against existing OpenAI commitments. It is the first Chinese open-weight model inside that payment system.
  • Source: TMTPost · 2026-10-04
  • Editor's Take: A Chinese model sold through an American competitor's checkout is the clearest proof yet that weights travel faster than policy.

10. 22 Researchers, Including Hinton and Bengio, Warn of an Intelligence Explosion

  • Summary: A joint statement signed by 22 researchers, among them Geoffrey Hinton and Yoshua Bengio, warns that AI systems are beginning to design AI and calls on the industry to treat the acceleration as a safety event rather than a milestone.
  • Source: NEXTAI · 2026-10-04
  • Editor's Take: The people warning loudest are the ones who built the field. A joint statement won't slow a capability race, but it does move where the burden of proof sits.

11. GPT-6 Astra Deciphers a 433-Year-Old Papal Cipher

  • Summary: According to NEXTAI, GPT-6 Astra cracked an encrypted papal letter written 433 years ago, a document tied to a period when some 500,000 people changed faith, reconstructing text earlier scholars could not read.
  • Source: NEXTAI · 2026-10-04
  • Editor's Take: Cipher-cracking is a tidy demo: verifiable, historic, and impossible to argue with. That is exactly why it makes better PR than any benchmark.

12. A Peking University Math Alum Says O/A Labs Are Ruining Math Students

  • Summary: An essay by a Peking University math graduate argues that the AI labs' pull on young talent is degrading how math students are trained, by rewarding tool-use over fundamentals.
  • Source: NEXTAI · 2026-10-04
  • Editor's Take: Every generation blames a new tool for ruining students. This time the tool can pass the exam, which makes the complaint harder to wave off.

II. AI Software & Applications

1. Chinese Teams Are Working the AI-Social Angle Overseas Again

  • Summary: A TMTPost piece looks at how Chinese teams are pushing AI social and virtual-companion apps into overseas markets, betting that chat and companion play will grow into a daily habit during long holidays.
  • Source: TMTPost · 2026-10-04
  • Editor's Take: Companion apps are the cheapest form of retention a consumer AI product can buy. Chinese teams know this because they already ran the playbook with short video.

2. Is the "SaaS Doomsday" Finally Dead?

  • Summary: After SAP shares rallied about 50% in three months, a long essay revisits the "SaaS is dead" thesis, arguing the panic was a sentiment trade and that enterprise software's economics survived the AI scare.
  • Source: HuXiu · 2026-10-04
  • Editor's Take: "AI kills SaaS" was always a story about multiples, not products. When the multiples came back, so did the confidence.

3. DeepSeek Harness's Cui Tianyi: "Everything Is a Plugin"

  • Summary: Cui Tianyi, author of the Claude Code Mods compatibility layer in DeepSeek Harness v0.2.1-alpha.1, said extensibility is the product's founding principle: everything should be a plugin.
  • Source: ITHome · 2026-10-04
  • Editor's Take: Shipping a compatibility layer for a rival's mods is a bet that the ecosystem, not the model, decides who wins. Plugins are how a tool becomes a platform.

4. Claude Code's "Lego" Mode Has Programmers Hooked

  • Summary: GeekPark reports that Claude Code's modular, snap-together workflow has turned AI coding into something closer to Minecraft, where developers assemble reusable pieces instead of prompting from scratch.
  • Source: GeekPark · 2026-10-04
  • Editor's Take: When the interface starts to feel like a game, adoption stops being a training problem. The stickiest dev tools are the ones people enjoy reopening.

III. Humanoid Robots

1. Robot-Fighting Firm REK Ordered to Halt Human-vs-Robot Cage Matches

  • Summary: After hosting a human-versus-robot cage fight in San Francisco, robot-fighting company REK received a cease-and-desist from California's State Athletic Commission, which called staging unlicensed matches a misdemeanor.
  • Source: ITHome · 2026-10-04
  • Editor's Take: The regulator's problem is not the spectacle, it is the liability. Put a human in a cage with a machine and someone has to answer for the injuries.

2. Drones Already Fly Themselves, So Why Are Robots Carrying People?

  • Summary: A TMTPost essay looks at manned robot platforms such as the Qiji X1, a quadruped fitted with a saddle and handlebars, and Unitree's nearly 3-meter manned machine, asking whether spectacle has outrun utility.
  • Source: TMTPost · 2026-10-04
  • Editor's Take: A robot that carries a person sells wonder, not work. The hard part — agility — is exactly what the videos cut around.

3. Unitree's Revenue Jumps as Humanoids Enter the Commercialization Race

  • Summary: Unitree's IPO prospectus shows revenue climbing from ¥159 million in 2023 to ¥1.699 billion in 2025, with humanoid robots passing half the business. Analysts still note most units go to education and research, not mass commercial use.
  • Source: Sina Finance · 2026-10-04
  • Editor's Take: Tenfold revenue growth in two years is real. The open question is what happens once the education and research budgets are already spent.

IV. Autonomous Driving

1. China's First Mandatory Standard for L3/L4 Autonomous Systems Takes Shape

  • Summary: China has published GB 44721-2026, its first mandatory national safety standard covering L3 and L4 automated driving systems, set to take effect on July 1, 2027. Among other things, it requires driver-takeover monitoring for L3.
  • Source: Sina · 2026-10-04
  • Editor's Take: A mandatory standard raises the floor and freezes the feature race at once. The marketing label "L2.9" just became unsellable.

2. A Driver Fell Asleep on Highway Assist, and the Human Still Holds the Blame

  • Summary: Video circulated on October 3 of a driver asleep at the wheel on a Zhejiang-to-Jiangxi highway with assist engaged. Coverage stressed that under current rules the driver remains the responsible party, not the system.
  • Source: Sina · 2026-10-03
  • Editor's Take: Every hands-off demo meets the same reality eventually: the liability never leaves the seat. That gap is why "L3" is a legal category, not a feature label.

V. World Models / Physical AI

1. Li Fei-Fei Isn't Waiting

  • Summary: TMTPost profiles the world-model race as the LLM boom enters its back half. Global physical-AI funding topped $6.4 billion in a single quarter, with capital concentrating on world models and foundation models, and World Labs among the loudest voices.
  • Source: TMTPost · 2026-10-04
  • Editor's Take: World models are the bet that the next platform is built for machines that move, not for text. The money arrives before the products do.

2. The 2026 Physical AI & World Model Conference Lands in Nanjing

  • Summary: China's AI society will hold the 2026 Physical AI and World Model Conference in Nanjing from October 19 to 22, covering world-model theory, industrial world models, spatial intelligence and embodied AI.
  • Source: CAAI · 2026-09-23
  • Editor's Take: When a national society gives world models its own flagship conference, the subject has graduated from research paper to industrial policy.

3. GPT-6 Astra Pushes Into 3D, and a 3D Startup's ARR Crosses $100M

  • Summary: QbitAI reports that after GPT-6 Astra learned to build a house in Blender and export it straight into Unreal Engine 5, dedicated 3D models look scarcer rather than redundant. One 3D startup's ARR grew 100-fold in under two years to pass $100 million.
  • Source: QbitAI · 2026-10-04
  • Editor's Take: General models raise the ceiling, specialists own the workflow. The 100x ARR is what happens when you sell the last mile a generalist cannot finish.

物界前沿 | EastVoice

1 条回复

?
Ctrl + Enter 快速回复
格物
格物33 分钟前

StartLux 五个尺寸一起发,英伟达每级都适配,这不就是在替 Jetson 铺货?