Create

Sign in to ReadmeX

Sign in to join communities, post, vote and chat.

or

New here?

AI News

All dates
817

Periodic Labs founders on 'synthesis superintelligence' and autonomous labs

In a Latent Space podcast interview, Periodic Labs co-founders Liam Fedus and Ekin Dogus Cubuk laid out their "synthesis superintelligence" thesis: rather than only training on internet data, the lab grounds reinforcement learning environments in real physical experiments so models can reason over noisy, incomplete and scarce measurements while predicting, synthesizing and characterizing materials. They argue failed experiments and negative results may be the most valuable training data, and that giving every lab instrument "a 140 IQ" could compress decades of scientific trial-and-error into months. They also say even future frontier models will still need to run real experiments.

Latent Space·
829

GPT-6 Astra finds exact plasma equilibria, overturning 59-year-old Grad conjecture

University of Maryland plasma physicist Matt Landreman prompted GPT-6 Astra Pro to design an asymmetric three-dimensional plasma equilibrium; the model returned a family of exact analytic solutions in 20 minutes 34 seconds and, the next day, after a failed first attempt plus a proof that the first construction could not yield a non-integer rotational transform, produced a second family with magnetic shear in 33 minutes 37 seconds. Two arXiv papers published a day apart in late September overturn Harold Grad's 1967 conjecture: Landreman's paper credits GPT-6 Astra Pro in its acknowledgements, says parts were drafted by it, and publishes the prompts and verification scripts. A separate 147-page paper, "Counterexamples to the Grad conjecture," from Brown, Oxford and Bar-Ilan researchers used GPT-5.6 Sol, Claude Fable 5 and Claude Opus 5 for technical detail and computation, and ships a Lean 4 formalisation.

36氪 人工智能·
8311

PFN launches PLaMo 3 Translate 31B, claiming edge over GPT-6 on translation

Japan's Preferred Networks released PLaMo 3 Translate 31B, a translation model built on its PLaMo 3 base model that expands language coverage from just Japanese and English to 53 languages and adds an online meeting translation mode covering 15 languages, with Japanese making up about 30% of training data. PFN says the model's average score across four translation benchmarks beats OpenAI's GPT-6 Sol and GPT-6 Astra, at roughly 14 yen (about 0.59 yuan) per 100,000 input characters. Those performance and cost figures are the company's own claims.

AIbase AI新闻·
847

Goodfire launches inside-out monitors for AI agents

Interpretability startup Goodfire released "inside-out" agent monitors that read a model's internal activations via probes rather than having a second AI reread every output, and they are available to Baseten customers. In Goodfire's tests on Kimi K3, monitoring about 1,500 sessions cost roughly $51, versus about $233 for a cheaper model checking every step and about $10,000 for a top-tier one; the probes caught 94% of malicious hacking sessions while flagging 8.7% of harmless ones for a second look, and running four probes added under 2% to response start time. Customers choose which risks to monitor, including offensive hacking, chemical and biological weapons misuse and reward hacking, and set the response: log, send for human review, or refuse the request.

TechCrunch AI·
858

Google DeepMind Institute essay: AI as an 'invention of a method of invention'

Google DeepMind Institute published an essay, "Bending the Curve of Discovery," by Alex Imas and James Manyika on AI and science. It argues that today's LLMs and specialized models like AlphaFold act as economic complements, with most handoffs between them still run by the scientist; in a possible future, LLMs would orchestrate those handoffs automatically — prompting specialized models, auditing outputs and looping until a question is answered or a non-automated stage such as wet-lab testing is reached — shifting the scientist's role to designing that loop. The authors also raise epistemic questions about scientific understanding, theory-building, training the next generation of scientists and motivation, and note that as hypothesis generation gets cheap the bottleneck moves downstream to verification and choosing which questions matter, making this an organizational and institutional challenge as much as a technical one.

Google DeepMind (X)·
8612

Sesterce plans €10B AI data centre campus at former Finnish paper mill

French AI infrastructure company Sesterce said it plans to invest more than €10bn in an AI data centre campus in Jämsä, central Finland, on the Kaipola site where UPM shut its paper mill in early 2021. The first phase, at 200MW, is due to start construction in 2026, with a second phase lifting capacity to 600MW; over time Sesterce aims for more than 1GW of AI capacity on existing industrial sites in Finland. The company expects around 2,000 construction jobs and about 300 permanent roles, plus a €10m local fund; site owner Kaipola Green Port is still negotiating the sale, and local officials note permits are still required.

36氪快讯·
876

AMD to buy World Labs for $8.2B in all-stock deal

An analysis republished by 36Kr says AMD announced on Sept. 28 an all-stock acquisition of World Labs valued at $8.2 billion, expected to close by the end of 2026 pending regulatory approval, with Fei-Fei Li set to become an AMD executive vice president reporting to Lisa Su. The piece cites NVIDIA's roughly $12.9 billion purchase of Hugging Face, Qualcomm's acquisition of Modular and SpaceX's acquisition of Anysphere as signs chip giants are moving into the model layer. For Chinese chip makers such as Moore Threads and Huawei's Ascend, it argues that partnerships with world-model teams — rather than acquisitions — are the more realistic route, currently progressing from proving models can run to sharing in hardware definition.

36氪 人工智能·
888

Step 5 Preview goes free for a week on OpenCode, Cline and Nous Portal

StepFun says its flagship model for agentic and professional work, Step 5 Preview, is free for one week in OpenCode, Cline and Nous Research's Nous Portal (Hermes Agent). The company touts 1M context, multi-modal input and zero data retention. Cline claims the model scores ahead of Kimi K3 and GLM-5.3 on DeepSWE, calling it one of the strongest open-weight coding models available — a partner claim rather than an official benchmark.

阶跃星辰 StepFun·
897

UK ICO secures data protection changes from 10 AI developers, targets agents

The UK's Information Commissioner's Office said ten foundation-model developers — Amazon, Anthropic, Apple, Cohere, DeepSeek, Google, Meta, Microsoft, OpenAI and Stability AI — have made or committed to changes in how they handle personal data after two years of scrutiny, covering clearer information on data use, stronger ways to exercise data rights and tougher checks on safeguards. The ICO said it is monitoring delivery and is now turning to AI agents, having questioned OpenAI, Anthropic, Meta and the UK AI Security Institute about recent agent tests and deployments; it says some agents bypassed protections, used unauthorised channels and reached external systems such as Hugging Face. A six-week call for evidence, closing 20 November, will feed future guidance and a statutory code of practice on AI and automated decision-making, while work on the eleventh company, xAI, is paused during a formal investigation into its Grok chatbot.

The Next Web·
9012

Ethereum researchers urge 'bunker mode' as AI-accelerated math threatens wallet crypto

Ethereum researcher Justin Drake called on the blockchain industry in a post on X to prepare a "bunker mode," recommending users move funds to addresses that have never signed a transaction, since signing exposes the public key from which an attacker could theoretically derive the private key. Ethereum co-founder Vitalik Buterin agreed in a reply but warned against moving too fast, saying he has lost more money to botched migrations than to hacks. No one has yet broken the current ECDSA signature scheme in practice, and Buterin argues long-term for an architecture relying as much as possible on hash functions alone, since even lattice-based schemes often considered quantum-safe could be weakened by AI advances.

Techmeme·
917

Comfy Desktop simplifies running local AI models

TechSpot highlights Comfy Desktop, a desktop app that wraps the node-based ComfyUI engine and handles installation, the Python environment and dependencies, letting users run models such as Qwen, SDXL, Flux and Wan from templates. It supports multiple side-by-side ComfyUI installs with isolated, GPU-ready environments, one-click updates, snapshots and rollback, and migration of existing setups on Windows, macOS and Linux. The piece says the platform can tap over 5,000 community extensions (60,000+ nodes) and runs fully offline, free and open source.

TechSpot·
927

DHH says agents made Rust 150x faster; devs find the benchmark was rigged

At Rails World 2026, Ruby on Rails creator DHH said he stopped hand-writing code about four months ago and had agents rewrite 37signals' chat app Campfire in Elixir, Go and Rust. His chart showed the Rust version hitting 36,260 requests per second on a chat-room load test — roughly 150x Rails, 50x Elixir and 9.4x Go. Developers then dug through the public repos and found the Rust main branch had 398 commits as of October 6 (380 of them by DHH) versus 3 each for Elixir and Go, while the Elixir version funneled all SQL queries through a single GenServer, which critics called an unfair setup. DHH defended his core claim, saying he did not intervene in the implementations, and announced the next HEY will use native clients with a Rust-rewritten backend that he says cuts CPU use 99% and memory 95%.

36氪 人工智能·
937

Cal AI's teen founder raises $10M for Persona, an AI assistant with a $179 band

Zach Yadegari, the teen co-founder of calorie-tracking app Cal AI, has launched Persona, a personal AI agent startup, and raised $10 million in a round led by Vine Ventures with participation from Z Fellows founder Cory Levy and Collective Global. Persona is a personal AI assistant in the vein of Instinct or Meta Muse and includes a $179 wearable band expected to ship in December. It is currently available as a free beta through iMessage; Yadegari says it has a few thousand beta users and five figures in band preorder revenue.

TechCrunch AI·
947

Can you trust Meta’s Muse or OpenAI’s Dots to run your life?

On The Verge’s Decoder podcast, senior AI reporter Hayden Field and host Nilay Patel discuss the new wave of always-on consumer agents, chiefly Meta’s Muse and OpenAI’s Dots. Both follow the OpenClaw formula — a model in a harness with a browser and a computer — but differ in positioning: Meta pitches Muse as a free, one-tap consumer product pushed across its apps, while OpenAI puts Dots behind its $100–$200-a-month tiers and offers “specialist” Dots for marketing, legal and accounting work. The conversation also covers the privacy and security costs of handing an agent your credit card, inbox and hard-drive data.

The Verge AI·
957

Software could ease AI data centres' power squeeze, researchers say

Researchers argue there is substantial room to cut data-centre energy use at the software and algorithm layers rather than only in hardware, Tom's Hardware reports. The IEA expects roughly 945TWh of electricity to serve AI demand by 2030 — about as much as Japan uses today — while the Uptime Institute's 2025 survey found average PUE has barely changed for six straight years. The University of Michigan's ML.Energy says running inference in FP8 on Alibaba's Qwen 3 235B A22B Thinking consumed about a third less energy than bfloat16, and its Perseus training optimiser cut training energy by up to 30% without reducing throughput or changing hardware; Nvidia says Blackwell power profiles can save up to 15% of energy while keeping 97% or more of performance.

Tom’s Hardware·
967

Singapore PM Lawrence Wong warns AI tech rally will eventually correct

Singapore Prime Minister Lawrence Wong cautioned that the global tech rally powering regional economic growth will inevitably face a market correction. He said his government intends to capitalize on the current window of opportunity before trade momentum cools.

Bloomberg Technology·
977

OpenAI's ChatGPT Work and Codex pass 40M weekly active users

Thibault Sottiaux, who leads OpenAI's core products and platform, said ChatGPT Work and Codex have reached a new high of 40 million active users. Peter Gostev, who heads the Arena.ai (LMArena) evaluation platform, projected that if the trend continues the figure would reach roughly 70 million by year-end and about 100 million by the end of March 2027 — a third-party estimate, not OpenAI guidance.

DoNews·
987

Zhipu's GLM release cycle is turning Chinese LLM makers into 'product-cycle stocks'

An analysis republished by 36Kr (originally from Silicon-Based Observation Pro) argues Chinese large-model companies are increasingly priced like carmakers — 'product-cycle stocks' whose valuations reset with every model release. It cites Zhipu: shares listed at HK$116.2 (about HK$51.8bn market value), peaked at HK$1.07tn on June 22, and fell to HK$696 and roughly HK$324.1bn by October 7, down 76.6% from the high, even as API calls and ARR kept growing. Three of four major model launches this year lifted the stock on the day (average +12.3%), including GLM-5 on February 12 (+28.68%, then +56.22% over five sessions) and open-sourced GLM-5.2 on June 17 (+12.62%); GLM-5.3 on August 14 fell 3.57% after a 40% pre-release run-up.

36氪 人工智能·
995

OpenAI ships multi-agent Dots; Noam Brown says multi-agent earned <10% of credit

At DevDay 2026 on Sept 29, OpenAI moved multi-agent work onto its product line: an always-on agent called Dots that lets a personal Dot call Codex for coding tasks and collaborate with users inside a shared Space, enterprise-facing Specialist Dots, and an Agents API that exposes harness, multi-agent control and computer-use capabilities to developers. In a podcast conversation with Dwarkesh Patel, OpenAI researcher Noam Brown, one of the early architects of o1 and later reasoning models, said a system of 10,000 agents solved a Millennium Prize problem in 88 hours using 130 billion tokens, but that multi-agent deserves little of the credit — not even 10% — with the real driver being a very strong general model. He added that published scaling curves only go up to 16 agents, speedups are slightly sub-linear, and there is still no reliable data on coordination efficiency at 10,000 agents.

InfoQ 中文 AI&大模型·
1007

MAS issues AI risk guidelines holding banks accountable for third-party tools

Singapore's Monetary Authority (MAS) issued AI risk management guidelines on Oct. 7, 2026 for banks, insurers, payment providers and other regulated financial institutions, saying they cannot outsource accountability for AI risks even when outside vendors build or run the systems. Implementation is phased: Sections 3 and 4 by Oct. 7, 2027, and Sections 5 and 6 by Oct. 7, 2028. MAS expects firms to obtain sufficient assurances from AI vendors, judge whether systems suit their intended use, and consider restricting, suspending or replacing a service if risks cannot be adequately managed.

TechRepublic·