Create

Sign in to ReadmeX

Sign in to join communities, post, vote and chat.

New here?

AI News

All dates
815

Claude Code’s Suggested Messages May Serve the Model First

A blog post discusses Claude Code’s suggested message feature and argues that the model, rather than the human user, may be its real customer. The item was surfaced through comments on Hacker News.

Hacker News · AI(100+ 分)·
823

OpenAI’s 28-Day Push Starts With GPT-6 Speed Claims and User Skepticism

OpenAI reportedly began a 28-day Codex and Work improvement push by increasing the default inference speed of GPT-6 Astra and GPT-6.1 Sol by about 50%, according to Tibo and coverage of the announcement. The report says the rollout drew skepticism because user tests allegedly fell short of the claimed 50 TPS and coincided with reports of ChatGPT visual ads, EU text watermarking, and changes to subscription value.

量子位(原生 RSS)·
835

Vast pitches tiered storage to ease AI agent memory pressure

In an interview with theCUBE at CoreWeave's Fully Connected 2026 event, Vast Data co-founder and CTO Alon Horev said agent memory differs from ordinary inference: it includes both in-session context and long-term memory that lets an agent review past interactions. Vast's approach tiers storage — GPU memory first, then CPU memory on the same machine, then persistent media holding petabytes of KV cache — with Nvidia's Dynamo software orchestrating the process. Horev said a 500,000-token session can occupy one-tenth to one-twentieth of a GPU's memory, and offloading such sessions to storage avoids repeat recalculation; enterprises also need to record and retain everything their agents do, data that can feed fine-tuning or purpose-built models.

SiliconANGLE AI·
845

Mirror Particle pitches a world model for changing human behavior

Mirror Particle is developing a foundation model intended to simulate how human behavior changes over time, using visual, social and other data alongside language. The San Francisco startup says its system focuses on revealed behavior and is initially targeting market research and brand strategy, while reporting that it has raised an angel round and is close to closing its first venture round.

TechCrunch AI·
853

Why US communities are resisting AI data centers

The article examines growing US community opposition to AI data centers, citing concerns over electricity prices, water use, noise, limited long-term employment and strained local infrastructure. It argues that fragmented power markets, lengthy permitting processes and the rush to expand AI capacity are deepening tensions between technology companies, governments and residents.

36氪 人工智能·
865

Microsoft Research podcast: what AI evaluation gets wrong

In a Microsoft Research podcast episode, host Chad Atalla speaks with Jennifer Neville, who leads the AI Interaction and Learning team at Microsoft Research and is a professor at Purdue, about how evaluation pushes the performance boundaries of today's AI systems and about the “surprising failures” that appear when models are tested beyond traditional benchmarks. Neville argues that the benchmarks commonly used in ML/AI are fairly simple relative to real-world use, so her team designs evaluations around multiturn behavior, collaborative settings and long-horizon tasks to expose performance gaps and then drive algorithmic and model improvements. She also offers practical guidance for working with current AI systems and stresses examining the data closely when results defy expectations.

Microsoft Research Blog·
874

AWS explains ISO/IEC 42005:2025 AI impact assessment guidance

AWS published a post on how organizations can use the ISO/IEC 42005:2025 standard to run AI system impact assessments and fold them into existing enterprise risk, privacy and security reviews. It covers the standard's guidance on when to assess, lightweight triage, required documentation and reassessment triggers, plus Annex D for integrating existing assessments without duplication and Annex E for a standalone template. AWS also points to its Well-Architected Responsible AI Lens and its ISO/IEC 42001 implementation guide, and says Amazon Bedrock, Amazon Q Business, Amazon Transcribe and Amazon Textract hold ISO/IEC 42001 certification.

AWS Machine Learning Blog·
884

Best practices for Amazon SageMaker HyperPod administration and governance

An AWS Machine Learning Blog post explains how to administer Amazon SageMaker HyperPod clusters through Amazon SageMaker Unified Studio while preserving underlying governance controls. It lays out four layers of control — organization, project, cluster, and workload — and covers designing identity, capacity, and observability policies across them, plus a "connection contract" record for each approved project-to-cluster connection.

AWS Machine Learning Blog·
894

SageMaker Studio can now manage HyperPod Spaces without the CLI

AWS says users can now create, configure, start, stop and open Amazon SageMaker Spaces on Amazon SageMaker HyperPod EKS clusters directly from the SageMaker Studio UI, instead of relying on the HyperPod CLI or kubectl. A new IDE and Notebooks tab on the cluster detail page offers a guided form and a searchable Spaces table, with browser access to JupyterLab or Code Editor and remote VS Code access over SSH-over-SSM. AWS says Karpenter over-provisioning can cut Space startup from 5–7 minutes to roughly 30–40 seconds.

AWS Machine Learning Blog·
904

Mitra adds chat takeover and Mitra Link to its AI assistant

Mitra is adding two capabilities to its personal AI assistant: chat takeover, which lets it reply in a thread from the user's own account and own an entire workstream across connected apps, and Mitra Link, a link-based format that lets one person's Mitra talk directly to other people's Mitras, with the user controlling who can reach theirs. Company posts cite real-estate agents coordinating showings agent-to-agent and consultants setting up client kickoffs, though those examples come from Mitra's own promotion and are unverified. Mitra previously focused on making and answering calls from a user's number, and the relaunched platform also handles texts, email, Slack, appointments, documents, spreadsheets and web tasks.

TestingCatalog·
913

Utopai X Reportedly Ranks Second in Artificial Analysis Video Test

Utopai Studios’ Utopai X reportedly scored 1,150 (±10) in an Artificial Analysis blind video evaluation, ranking second globally and first in the United States. The article says the model was post-trained from MiniMax H3 for film production and led the evaluation’s audio-sync and physics categories, but these claims are presented through a promotional report and are not independently verified here.

新智元·
923

OpenAI Brakes on Agent Safety as Meta Keeps Pushing

A Hu Xiu AI analysis contrasts OpenAI’s reported decision to pause training, evaluation and tool use for an advanced agent after internal sandbox incidents with Meta CEO Mark Zuckerberg’s opposition to collectively slowing AI development. The article argues that both pre-release safety gates and post-release market or legal accountability remain incomplete, citing reported vulnerabilities affecting Meta’s Muse and broader concerns about agent autonomy.

虎嗅 AI·
933

Why Stronger AI Has Not Made Companies Deliver Faster

An analysis argues that faster AI-generated outputs do not automatically make companies deliver work faster. Teams still need clear distinctions between demos, usable results, and accepted deliverables, along with processes for context gathering, verification, review, and decision-making.

虎嗅 AI·
944

NetApp pitches AI-ready legacy data without a rebuild; plans Oracle Cloud service

At its NetApp INSIGHT event, NetApp laid out an "intelligent data infrastructure" strategy that it says lets enterprises make legacy data usable for AI in place, across file, block and object storage, without re-architecting or moving it. SVP of product marketing Jen Prenner said data must be accessible, governed, protected and available, adding that protection is built into every layer, that the company has expanded its Commvault partnership, and that its Console control plane can run locally for disconnected environments. NetApp also announced plans for a fully managed storage service on Oracle Cloud Infrastructure within the next 12 months.

SiliconANGLE AI·
953

Why AI Should Move Toward “Democratized Innovation”

The article argues that AI’s rapid technical progress has not translated into broad productivity gains, and that current innovation remains concentrated on automating or replacing labor. Drawing on the “productivity paradox,” “Polanyi’s paradox,” and “innovation paradox,” it calls for open-source tools, shared platforms, public support, and worker participation to steer AI toward more inclusive applications.

虎嗅 AI·
964

A 2026 guide to five robot companions

TechRepublic compares five robot companions across interaction style, mobility, privacy, battery life, and price: Casio Moflin, Loona Petbot, Mirumi, Ropet KAMOMO, and Eilik. The roundup recommends Moflin for adaptive, pet-like interaction, Loona for movement and connected features, Mirumi for portability, KAMOMO for customization and offline AI, and Eilik as the least expensive option.

TechRepublic·
974

Plaud, Omi and More: Wearable AI Note Takers Compared

A TechRepublic comparison examines four wearable AI recorders: Plaud NotePin S, SwitchBot AI MindClip, Omi and Vocci AI Ring. It recommends Plaud for overall flexibility, SwitchBot for value, Omi for open-source integrations and Vocci for a ring form factor, while noting that several battery figures come from manufacturers.

TechRepublic·
984

Ben's Bites: four kinds of AI users, and why personal agents aren't there yet

In a set of reflections from a week in San Francisco, Ben's Bites sorts AI users into roughly four groups: everyday people treating it as a better Google, curious non-engineers who try Lovable or Replit but quit at the first database bug, non-engineers who can steer a coding agent, and professional developers. The author argues personal agents will be the interface of the future but aren't yet, citing three pet peeves: one thread for everything, invisible memory, and loops that never close across apps. The piece also notes briefly that Claude Code can now be modded like a video game and that Google Docs opens Markdown files directly.

Ben's Bites·
993

A16Z reports: AI usage is broad, but deep adoption remains concentrated

An analysis of two Andreessen Horowitz reports argues that AI adoption is broader than its paid and intensive use would suggest. It highlights a widening gap between casual users and power users, falling token costs driving greater demand for agents, and AI infrastructure spending flowing into chips, energy, and other physical assets.

虎嗅 AI·
1003

China’s 1990s-born AI leaders take the helm as the real test begins

A 36Kr feature examines a group of Chinese AI leaders born in the 1990s who now oversee major model, multimodal, agent and AI infrastructure efforts at Alibaba, Xiaomi, ByteDance and Tencent. It argues that their rapid promotions reflect the shorter cycle of the foundation-model industry, while emphasizing that technical performance, commercialization, cost control and organizational execution remain unproven.

36氪 人工智能·