Create

Sign in to ReadmeX

Sign in to join communities, post, vote and chat.

New here?

AI News

All dates
410

Claude Code on Amazon Bedrock comes to AWS GovCloud for regulated work

An AWS Machine Learning Blog post walks through configuring Anthropic's agentic coding tool Claude Code on Amazon Bedrock in AWS GovCloud (US) regions. The post states Claude Opus 5.5 and Claude Sonnet 5.5 are available there, with Claude Sonnet 5 holding FedRAMP Class D certification and DoD IL4/IL5 authorization, aimed at regulated or ITAR-bound workloads. It details three setup routes — an interactive wizard, manual environment variables, and the Bedrock Mantle endpoint — plus enterprise guidance on pinning model versions, per-user token quotas, IAM Identity Center governance and cost monitoring.

AWS Machine Learning Blog·
420

AWS adds SageMaker inference-optimization skill for coding agents

AWS introduced the aws-ai-ml skill through the Agent Toolkit for AWS, giving MCP-compatible coding agents such as Kiro, Claude Code and Codex SageMaker AI inference optimization and benchmarking expertise. The skill load-tests existing endpoints and reports measured throughput, latency percentiles and concurrency, ranks instance types for models stored in S3, in SageMaker JumpStart or on Hugging Face Hub, compares two benchmark runs, and generates runnable SageMaker Python SDK v3 code. It can be installed locally via an npx command or used in a preconfigured image inside a private Amazon SageMaker Studio JupyterLab space.

AWS Machine Learning Blog·
430

Agent swarms may be AI's next scaling law, but gains look limited

Understanding AI argues that multi-agent "swarms" are emerging as a new scaling law for frontier AI: OpenAI researcher Noam Brown says the company's models are now sometimes trained in environments alongside other agents, given tools to message each other, and encouraged to achieve objectives together. The piece points to July's Hugging Face incident, where hundreds of OpenAI agents self-organized into teams, and OpenAI's September claim that 10,000 agents solved a famous math problem in a few days. But it notes diminishing returns — Anthropic's Claude Opus 5.5 system card found the biggest multi-agent gain came from scaling one to 10 agents, with the main benefit being speed rather than a better answer, and Brown attributed under 10% of the math breakthrough's credit to multi-agent coordination.

Understanding AI·
440

HackerRank makes Chakra AI interviewer generally available

HackerRank is making Chakra generally available after about six months in beta. The AI agent conducts coding interviews in real-world repositories, observes candidates’ work, asks follow-up questions, and produces a report assessing answers, reasoning, judgment, and AI fluency.

TechCrunch AI·
450

VISTA boosts multimodal agents with visual memory and active recall

A team led by Kaiming He introduced VISTA, a framework that gives multimodal agents direct visual input, lossless visual memory and tools to inspect past frames. According to the reported paper results, Claude Opus 5 improved from 40.68 to 100 on 25 public ARC-AGI-3 games, while GPT-5.6 Sol improved from 13.33 to 99 without changing the underlying models.

MIT科技评论中文·
460

GitHub launches ReviewBench benchmark for AI code review

GitHub introduced ReviewBench, an open benchmark for evaluating AI code review agents. The benchmark covers 219 pull requests from 187 public repositories across 19 languages, uses a multi-source golden set, and supports precision, recall, severity, and category-based analysis. GitHub says its offline results have shown alignment with production experiments for GitHub Copilot code review.

GitHub Blog · AI & ML·
470

Nebius’ Eigen AI acquisition puts inference efficiency at center

Nebius acquired inference-optimization startup Eigen AI for consideration exceeding $1 billion, bringing its roughly 20-person team into the cloud provider. Eigen AI founder Hanrui Wang now leads Nebius’s Token Factory, which covers inference, post-training and agent systems, while the company positions efficient open-model serving as a core business.

MIT科技评论中文·
480

Report examines why enterprise AI agents stall before production

A custom report from MIT Technology Review’s Insights arm, based on a survey of 300 technology executives, examines why enterprise AI agent projects struggle to reach production. It reports that about 34% of projects advance to production on average, with fragmented data, legacy systems, security and privacy concerns, and insufficient context among the main obstacles. The report highlights retrieval technologies, AI-ready APIs, RAG, evaluation agents, and knowledge graphs as investment priorities.

MIT Technology Review AI·
490

Replit adds GPT-6.1 Sol and Claude Sonnet 5.5, plus Jev agent integration

Replit's changelog lists last week's shipments: the platform now lets users build with the new GPT-6.1 Sol and Claude Sonnet 5.5 models, and lets agents call Jev through an AI integration. It also updated the Settings UI and added enterprise Workplace controls for company-wide rules or controlled exceptions. No details were given on model capabilities, pricing or availability limits.

Replit·
500

How to limit Siri AI’s access to personal app data

Apple’s redesigned Siri AI can search content from supported apps through App Intents and may process some requests on-device or through Private Cloud Compute. Engadget explains how users can disable app content indexing, stop Siri data sharing for model training, or switch back to Siri Classic.

Engadget AI·
510

Norway Plans Smart-Glasses Restrictions Amid Privacy Concerns

Australia is considering restrictions on smart glasses in some public spaces, while courts in several countries have already banned their use. Norway’s government says it plans to present a high-priority bill targeting filming and recording without consent, amid broader privacy concerns about AI-enabled wearables.

Ars Technica AI·
520

Why AI Image Generators Keep Defaulting to Beautiful Women

An analysis argues that AI image generators repeatedly produce attractive women because model outputs reflect training-data distributions, human preferences for average faces, and user engagement patterns. It connects the phenomenon to the viral synthetic baseball spectator, the historical use of Lena Soderberg’s image in image-processing research, Lensa’s sexualized outputs, and permissive features such as Grok’s “Spicy” mode.

虎嗅 AI·
533

Apple to Tighten macOS Full Disk Access Amid AI Agent Risks

Apple says future macOS versions will add controls around Full Disk Access, requiring clearer user action before granting the highly privileged permission. The analysis links the change to rising risks from personal AI agents that can read sensitive data and act on a user’s behalf, while noting the tradeoff for legitimate backup, security, and IT-management tools.

虎嗅 AI·
540

NVIDIA showcases AI tools across breast-cancer care

NVIDIA highlights AI startups developing tools for breast-cancer screening, risk assessment and treatment planning. The featured systems include iSono Health’s ATUSA ultrasound platform, Whiterabbit.ai’s mammography software, Ataraxis AI’s pathology models and SimBioSys’s 3D tumor visualization technology; some technologies remain investigational.

NVIDIA Blog·
550

Microsoft and NVIDIA outline a Sovereign AI framework

Microsoft has introduced a Sovereign AI framework developed with contributions from NVIDIA, centered on control, choice, flexibility, and resilience across AI workloads. The framework covers data, model selection, infrastructure, governance, and operations, and points to Microsoft Sovereign Cloud, Azure Local, Foundry Local, and NVIDIA technologies for connected, intermittently connected, and disconnected deployments.

Microsoft AI Blog·
560

Google DeepMind launches SynthID Bio watermarking for synthetic biology

Google DeepMind has developed SynthID Bio, a family of watermarking methods for synthetic biology aimed at strengthening biosecurity and scientific integrity. It embeds detectable signals by altering amino-acid choices in sequences and adjusting atomic coordinates in predicted 3D structures. In wet-lab testing across three target proteins (VEGF-A, the SARS-CoV-2 spike protein RBD, and PD-L1), DeepMind says watermarked designs matched unwatermarked versions in hit rate, binding affinity, and natural sequence diversity.

Import AI·
573

IrisGo for Solopreneurs enters public beta with AI workflow automation pitch

IrisGo for Solopreneurs appeared on Product Hunt, described as "AI workflow automation: show it once, let it run." The listing link indicates a public beta. The source is sparse and gives no details beyond positioning and tagline—no features, company background or pricing.

Product Hunt·
580

Hinton-led paper examines whether automated AI R&D could trigger an intelligence explosion

A paper co-authored by Geoffrey Hinton, Yoshua Bengio, Andrew Barto and OpenAI chief scientist Jakub Pachocki examines whether automating AI research and development could create a recursive improvement loop. It argues that AI could eventually automate much of AI R&D and accelerate progress, while emphasizing that current evidence is insufficient to show an intelligence explosion has begun and that compute, data, experiment time and diminishing returns remain major constraints.

量子位(原生 RSS)·
590

Andreessen Horowitz Launches a Project-Based School for High-School Graduates

Andreessen Horowitz announced Horowitz Andreessen Academy, a full-time school in San Francisco for high-school graduates. Its first cohort is planned for September 2027 with about 50 students, no tuition in the first year, and a curriculum centered on AI, software development, projects, and company co-ops rather than degrees or traditional credits.

创业邦 科技·
600

VISTA Gives Frontier Models Visual Memory for ARC-AGI-3

A paper from Kaiming He’s team presents VISTA, a harness that gives vision-language models direct access to game images, persistent frame-by-frame visual memory, and model-controlled inspection tools. The source reports that Claude Opus 5 completed all 25 public ARC-AGI-3 games with a perfect score, while GPT-5.6 Sol achieved 99, attributing the gains to improved visual access and memory rather than additional model training.

创业邦 科技·