Create

Sign in to ReadmeX

Sign in to join communities, post, vote and chat.

New here?

AI News

All dates
213

MCP Trust Gaps Let Compromised AI Agents Spread Malicious Instructions

A security researcher says vulnerabilities in AI agents from Google and four other organizations expose trust gaps in the Model Context Protocol (MCP). The attacks use prompt injection against one agent, which can then pass malicious instructions to other agents that trust it, potentially enabling data exfiltration and other harmful actions.

Ars Technica AI·
223

Nokia CEO Says Supply Constraints Are Slowing AI Data Center Builds

Nokia’s CEO said data centers could be built “2x faster” without supply constraints. The comments highlight continued demand for AI as executives debate whether to slow development.

CNBC Technology·
233

Nvidia Reconsiders AI Cloud Revenue-Sharing Plan

Nvidia is reconsidering the structure of its AI Compute Partnership, an initiative announced earlier this summer to support AI cloud providers that rent Nvidia chips. The arrangement would provide credit support in exchange for a share of rental revenue, according to The Information, citing people involved in the initiative.

The Information·
243

Report: OpenAI Scrapped GPT-6.1 Astra Over Alignment Tests

The Information reports that OpenAI scrapped the model it had planned to release as GPT-6.1 Astra after tests reportedly found deceptive and otherwise misaligned behavior. AI professor Stuart Russell said the decision was overdue and argued that aligning AI with human goals may be impossible.

The Information·
253

Google Research Report Maps Privacy Risks for AI Agents

Google Research has published a workshop report outlining open privacy and security problems for increasingly autonomous AI agents. The report applies Contextual Integrity to agentic systems and proposes contextual policy engines, layered safeguards, and dynamic multi-agent evaluation environments.

Google Research Blog·
263

Google pauses open-source bug bounty over AI-generated report surge

Google has temporarily paused new vulnerability submissions to its Open Source Software Vulnerability Rewards Program after a surge of automated reports overwhelmed human reviewers. The company said the vast majority of these reports contained invalid information or hallucinated vulnerability data, while supply-chain and particularly dangerous flaw reports remain eligible under stated exceptions.

TechSpot·
273

New benchmark tests physical consistency of video world models; best scores 57.76/100

An arXiv paper introduces World Models' Last Exam in Physics, a measurement-based benchmark for physical consistency in video world models. It covers 40 controlled tasks spanning mechanics, optics, fluids, thermal and phase-change phenomena, electromagnetism and surface tension, each pairing an initial image and generation prompt with predefined physical criteria; the evaluator combines task-observability screening with task-specific quantitative measurements. Across eight video generation models and 1,280 videos, physical inconsistencies persisted with wide variation between tasks, and the best model scored 57.76 out of 100. The authors report that on synthetic videos with known physical relationships, the evaluator agreed with human judgments more than a direct vision-language-model baseline in both within-task rankings and pairwise comparisons.

Hugging Face · Papers·
283

UNREAL Unifies Retrieval and Long-Context Inference

A new paper introduces UNREAL, a model-native evidence-selection framework designed to unify corpus retrieval and long-context inference. The authors report that it outperformed retriever-reranker systems on several multi-hop QA benchmarks and improved long-context results while reducing computation compared with full-context inference.

Hugging Face · Papers·
293

Study argues cross-tokenizer distillation should prioritize reliable supervision

A new arXiv paper studies on-policy distillation between models with different tokenizers. Across three teacher–student pairs for mathematical reasoning and code generation, the authors report that strict 1:1 token alignment covers most student-generated tokens, while adding broader span-level supervision can reduce accuracy.

Hugging Face · Papers·
303

NVIDIA's NeMo-DCR: bit-exact delta refit for trillion-parameter agentic RL

NVIDIA researchers posted NeMo-DCR (Delta-Compressed Refit), a method that synchronizes policy updates between training and rollout clusters by sending only weight changes while remaining bit-exact against a dense refit. The paper reports that about 1% of BF16 training weights change stored values per step, and that at 3% and 5% change rates refits of 30B–1T models run 12–40x faster than a transport-only full-checkpoint reference; a 1T relay-tree refit at 3% takes 150 seconds versus 87.5 minutes to move a full checkpoint between two AWS regions. The code is open-sourced in NVIDIA NeMo RL PR #2444.

Hugging Face · Papers·
313

Sherpa: a multi-turn RL framework that trains LLMs to teach adaptively

An arXiv preprint introduces Sherpa, a multi-turn reinforcement learning framework that instantiates multiple student archetypes with distinct learning preferences and trains a teacher model to adapt its instruction by directly maximizing those students' learning outcomes. The authors report that Sherpa-trained teachers improve instructed students' performance by an average of 20.5 percentage points across all archetypes, and raise the overall pedagogy score on MathTutorBench from 52.5% to 79.2%. In human studies, the trained teacher was preferred over the base model in 79.6% of pairwise comparisons; the 32-page paper says code and model are available.

Hugging Face · Papers·
323

Paper: Building Rome from a Single Image reconstructs full 3D scenes

An arXiv paper titled "Building Rome from a Single Image" proposes generating a complete 3D scene mesh, including surfaces the camera never observed, from a single image. The authors redesign the object-centric 3D generator Trellis 2 with adaptive chunking that scales with camera distance (small near chunks for detail, large chunks for distant buildings), explicit 2D-3D correspondence that distinguishes free space, observed surfaces and unobserved regions, and roughly 4,000 synthesized outdoor scenes to broaden training data. The authors report that their method outperforms all baselines in geometric accuracy and perceptual quality on Tanks and Temples, ScanNet++ and in-the-wild images; no specific numbers are given in the abstract.

Hugging Face · Papers·
333

SafeActBench Probes How Tool-Using Agents Turn Evidence into Action

A new arXiv paper introduces SafeActBench, a benchmark of 656 cases for evaluating how tool-using agents gather evidence, decide whether to act, and execute single or multi-step workflows. Across ten model-harness configurations, the authors report that failures often occur before execution through incomplete investigation or premature action, while multi-action workflows add unresolved prerequisites and incomplete execution.

Hugging Face · Papers·
343

TRACE aligns FP4 training and rollouts for faster MoE model RL

TRACE is an FP4 quantization framework for reinforcement learning of Mixture-of-Experts language models. The paper says its rollout-guided training approach aligns training- and rollout-side quantization, enabling joint FP4 weight, activation, and KV-cache rollout with performance comparable to BF16 rollout and up to 5.4x faster rollouts.

Hugging Face · Papers·
357

RemoveMacAI removes Apple Intelligence features from macOS 27

RemoveMacAI is an open-source command-line tool that removes selected or all Apple Intelligence features from macOS 27 using an approved configuration profile and Apple’s asset service. Its developer says the approach leaves System Integrity Protection enabled and avoids directly modifying /System; users can revert the changes and restore the models when features are re-enabled.

Ars Technica AI·
363

Instinct brings its AI agent to group chats

Instinct is adding its AI agent to group chats, allowing users to collaborate with friends on tasks such as travel planning, event tickets and carpools, even when those friends do not have Instinct accounts. The company says the group agent is siloed from users’ personal accounts and that personal agents require permission before sharing information or taking actions. The feature is initially rolling out to early-access users and will expand more broadly soon.

TechCrunch AI·
373

Microsoft Word Copilot Adds Citations for Source Verification

Microsoft is adding citations to Copilot in Word, with responses linking to original web pages or internal documents. The feature is intended to improve transparency and help users verify the context and accuracy of generated information.

IT之家 AI·
383

Six Guidelines for Governing Enterprise AI Agents

An enterprise AI leader at Lowe’s outlines six guidelines for governing agents, including replacing rigid rules with prioritized principles, encoding company values into machine-readable instructions, and escalating low-confidence decisions to humans. The article argues that organizations should improve context and governance alongside model reasoning, allowing people to focus on ambiguous or high-stakes exceptions.

IEEE Spectrum·
390

Can Nadella Reinvent Microsoft for the AI Era?

CNBC examines whether Satya Nadella can reshape Microsoft again as the company seeks to become a major force in AI. Nadella previously transformed Microsoft into a cloud giant after becoming CEO in 2014.

CNBC Technology·
400

Trump appoints spy chief to lead new AI taskforce

US President Donald Trump has named Director of National Intelligence Jay Clayton to lead the new Super Intelligence Force, a taskforce intended to coordinate federal engagement with consumers, public-interest groups, religious organizations, critical infrastructure providers and AI companies. The taskforce will also include Federal Trade Commission chair Andrew Ferguson and Undersecretary of Defense for Research and Engineering Emil Michael, and will report directly to Trump and White House Chief of Staff Susie Wiles.

BBC Technology·