NVIDIA
AI overview
Sign in and the AI will write an overview from our coverage.
Headlines · 30
- TechCrunch Disrupt 2026 roundtable lineup spans physical AI, agentic enterprise
TechCrunch published the full lineup of interactive roundtables for Disrupt 2026, which it says will bring 10,000+ founders, investors and operators to Moscone West in San Francisco on October 13-15. Sessions cover physical AI data challenges, moving enterprise agentic AI from pilot to production, AI's impact on SaaS, specialty post-training data, quantum computing and generative AI in healthcare.
TechCrunch AI · 🔥 9 - Nous Research raises $90M at reported $1.2B–$1.5B valuation for enterprise Hermes
Nous Research has raised $90M to build Hermes for Businesses, a commercial version of its open-source Hermes agent, with investors including Nvidia, Microsoft's M12, Samsung, Robot Ventures, Union Square Ventures, Y Combinator and Menlo Ventures, CEO Dillon Rolnick wrote. Nous did not disclose a valuation: WSJ, which first reported the news, was cited by Techmeme at $1.2B while The Next Web said $1.5B. Hermes was released under the MIT licence in February and has been downloaded more than 22 million times since, according to WSJ, while Nous says it has been cloned over 24 million times and estimates internally that it drives about 2.5% of global token usage.
Techmeme · 🔥 19 - Google opens SynthID Detector website to all users globally
Google said its SynthID Detector portal at synthid.com is now available to everyone worldwide, after more than a year of limited access for journalists and other testers. The free tool checks images, video and audio for SynthID watermarks from Google, OpenAI, NVIDIA and Kakao, with Apple support "coming soon," and requires signing in with a Google, OpenAI or Apple account.
Engadget AI · 🔥 66 - TechCrunch Disrupt 2026 opens in 6 days; online ticket discounts end Oct 13
TechCrunch Disrupt 2026 runs October 13-15 at Moscone West in San Francisco, with the organizer expecting 10,000+ people from the global startup and tech ecosystem. Online registration before doors open saves up to $100 and gets 50% off a second eligible pass, while laid-off attendees can buy a $75 Expo+ Pass. The event lists 200+ sessions across six stages, 250+ speakers, and a Startup Battlefield 200 where 20 finalists compete for a $100,000 prize.
TechCrunch AI · 🔥 9 - Biohub, Meta, Google DeepMind and US agencies commit $1.8bn to AI biology data
Biohub, the nonprofit research institute backed by Mark Zuckerberg and Priscilla Chan, announced a $1.8bn pooled effort to build open biology datasets for training AI models. Meta, Google DeepMind and Isomorphic Labs are contributing $300M combined; the US Department of Energy will spend more than $500M over five years through its Genesis Mission, the NIH is contributing datasets built with over $500M in earlier federal funding, Biohub itself has pledged $500M and Nvidia will supply computing and software. Commercial funders get one year of exclusive access before the data becomes public, government-funded work carries no such restriction, and partners aim for a first dataset in about a year and accurate predictive models within five years.
Techmeme · 🔥 28 - DeepGEMM adds Ascend support and new kernels, trends on GitHub
DeepGEMM is DeepSeek's open-source, high-performance tensor-core kernel library that unifies many core LLM computation primitives — FP8/FP4/BF16 GEMMs, fused MoE with overlapped communication (Mega MoE), MQA scoring for the lightning indexer (including a sparse version) and HyperConnection (HC) — in a single CUDA codebase, with all kernels compiled at runtime via DeepJIT and no CUDA compilation at install time. According to the repository's changelog, DeepGEMM-Ascend became available on 2026.09.30 alongside new optimizations such as locality domain features, after earlier additions of the Sparse Indexer, Mega Gate, Mega mHC and MoE/Indexer optimizations. The repo also says its performance matches or exceeds expert-tuned libraries across various matrix shapes, and that it reached up to 1550 TFLOPS on H800 in April 2025.
GitHub Trending(每日) · 🔥 26 - Nvidia Considered OpenRouter Deal Before Stripe’s $8B Bid
Nvidia explored a possible takeover of AI model marketplace OpenRouter but did not make a formal offer, according to a person familiar with the discussions. Stripe ultimately prevailed with an $8 billion bid, while Nvidia later pursued more than $140 billion in deals over two months, including a $105 billion credit guarantee.
The Information · 🔥 9 - Fine-Tuned Nemotron Reports Gold-Level IOI and IMO Results
A Hugging Face post reports gold-level results after fine-tuning Nemotron for the International Olympiad in Informatics (IOI) and International Mathematical Olympiad (IMO). The provided source contains no further details about the fine-tuning method or results.
Hugging Face · 🔥 9 - Microsoft and Nvidia to unveil Surface Laptop Ultra for local AI
Microsoft CEO Satya Nadella and Nvidia CEO Jensen Huang are expected to present the Surface Laptop Ultra in San Francisco on Wednesday. The laptop is built around Nvidia’s RTX Spark chip and is designed to run AI agents locally, but its price, release date and final availability have not been announced.
The Next Web · 🔥 9 - CoreWeave to build first India data centers with 240MW from AdaniConneX
CoreWeave said it will open its first India data centers, taking 240 megawatts at an AdaniConneX campus in Navi Mumbai and becoming the sole tenant across three 80-megawatt buildings, with an option for another 240 megawatts. The first phase is expected online in mid-2028, running Nvidia's Vera Rubin platform for training, inference, reasoning and AI agents, with direct liquid cooling for the GPU halls and a chilled-water system that does not rely on evaporation. Bloomberg reported the investment will total multiple billions of dollars over the project's life, though CoreWeave did not give a figure; this is its second Asia market after Indonesia, and the company will also open an India office and hire locally.
Bloomberg Technology · 🔥 18 - Foxconn’s AI-server business becomes a second growth engine
A report says Hon Hai Technology Group and Foxconn Industrial Internet are benefiting from surging AI-server demand, alongside their established Apple manufacturing business. Foxconn Industrial Internet reported strong 2025 revenue and profit growth, while its cloud-computing business and AI-server sales became major growth drivers, according to the cited financial disclosures.
钛媒体 · 🔥 7 - Caterpillar and CoreWeave team up to shorten physical AI learning loops
Speaking at CoreWeave's Fully Connected event, Brandon Hootman, Caterpillar's VP of physical AI platforms and construction autonomy, and Richard Ahlfeld, CoreWeave's SVP of physical AI, described their collaboration on autonomous construction equipment. Caterpillar's digital ecosystem already holds about 18 petabytes of federated data from machines, dealers and customers, while a single machine can generate terabytes of lidar, camera, control and performance data in a day. Working with Nvidia, the partners use AI models to annotate and label incoming field data, cutting work that used to take weeks or months down to hours so the simulation and training loop closes within a workday. CoreWeave named Caterpillar among its enterprise customers in its second-quarter results and launched a Physical AI Field Engineering service in September.
SiliconANGLE AI · 🔥 7 - SpaceX seeks $40B financing, led by Apollo, to buy Nvidia chips
SpaceX is seeking to raise about $40 billion to buy Nvidia chips in a financing led by Apollo Global Management, the Financial Times reported, with the deal expected to close in 2027 and consist of roughly $10 billion in bank loans and $30 billion in investment-grade bonds; Pimco is among the lenders in talks. Musk said previously that SpaceX has decided to build entirely on Nvidia technology, calling the Vera Rubin architecture the best AI platform. SpaceX, Nvidia, Apollo and Pimco did not immediately respond to requests for comment.
The Information · 🔥 23 - Anthropic Researcher Predicts Human-Surpassing AI Within Years
Anthropic reinforcement-learning lead Sholto Douglas said in an interview that a model capable of doing all computer-based work could emerge within a few years. He also projected that annual AI capital spending could reach $4 trillion by 2028 and that global GDP could double in the early 2030s, while colleague Nick Marwell warned of the risks once AI no longer needs human partners.
36氪 人工智能 · 🔥 6 - Dell adds knowledge graph, semantic layer and Knowledge Agents to AI Data Platform
Dell announced extensions to its AI Data Platform, adding a Unified Semantic Layer, an Enterprise Knowledge Graph and Knowledge Agents to its Data Orchestration Engine, plus a Data Processing Engine that runs on GPUs via NVIDIA's cuDF library. The platform ties storage (PowerScale, ObjectScale and the Lightning File System) to orchestration, governance, search and GPU acceleration to cut data movement and turn enterprise data into agent-ready context. Dell infrastructure chief Arthur Lewis said the constraint is often the data, not the model.
SiliconANGLE AI · 🔥 6 - Lambda reportedly seeks $4B ahead of planned 2027 IPO
AI cloud provider Lambda is reportedly raising up to $4 billion at a $14.5 billion pre-money valuation, potentially its final private round before a planned 2027 IPO. Coatue Management and Blackstone are leading the round, while much of Lambda’s reported $50 billion backlog appears tied to a $35 billion Anthropic commitment. The financing highlights both strong demand for scarce GPU capacity and the debt and customer-concentration risks facing neocloud providers.
TechCrunch AI · 🔥 6 - CoreWeave launches Forge to speed continuous AI post-training
CoreWeave announced CoreWeave Forge, which connects model deployment, evaluation and improvement into one loop; its reinforcement-learning Rollouts capability, now in preview, supports repeated cycles of generating training responses and updating models, with weight synchronization that hot-starts from nearby peers instead of pulling weights from object storage each time. CoreWeave AI Object Storage now supports cross-region writes so post-training jobs can write results back for others to use. You.com joined CoreWeave's partner network to give agents a search layer, and the companies say that working with NVIDIA they used RL Rollouts to post-train Nemotron 3.5 Lightning in eight hours with You.com's web search tools. You.com chief product officer Saurabh Sharma said the ceiling is no longer model intelligence but models' ability to use tools, and claimed customers get higher accuracy and lower total cost of ownership — a vendor claim without published figures.
SiliconANGLE AI · 🔥 6 - Vast pitches tiered storage to ease AI agent memory pressure
In an interview with theCUBE at CoreWeave's Fully Connected 2026 event, Vast Data co-founder and CTO Alon Horev said agent memory differs from ordinary inference: it includes both in-session context and long-term memory that lets an agent review past interactions. Vast's approach tiers storage — GPU memory first, then CPU memory on the same machine, then persistent media holding petabytes of KV cache — with Nvidia's Dynamo software orchestrating the process. Horev said a 500,000-token session can occupy one-tenth to one-twentieth of a GPU's memory, and offloading such sessions to storage avoids repeat recalculation; enterprises also need to record and retain everything their agents do, data that can feed fine-tuning or purpose-built models.
SiliconANGLE AI · 🔥 5 - NVIDIA promotes open models for telecom AI and announces Nemotron 3 LTM
NVIDIA says telecom operators are increasingly using open models to gain more control, customization and deployment flexibility across network operations and customer care. It also announced the 30-billion-parameter Nemotron 3 Large Telco Model, fine-tuned by AdaptKey on open telecom datasets, and released a NeMo-based recipe for adapting open models to operator-specific data.
NVIDIA Blog · 🔥 5 - OpenAI Brakes on Agent Safety as Meta Keeps Pushing
A Hu Xiu AI analysis contrasts OpenAI’s reported decision to pause training, evaluation and tool use for an advanced agent after internal sandbox incidents with Meta CEO Mark Zuckerberg’s opposition to collectively slowing AI development. The article argues that both pre-release safety gates and post-release market or legal accountability remain incomplete, citing reported vulnerabilities affecting Meta’s Muse and broader concerns about agent autonomy.
虎嗅 AI · 🔥 4 - Ghost AI raises $11M to build Core, a $3,499 personal AI agent computer
AI hardware startup Ghost AI said it raised $11 million led by Andreessen Horowitz, with Abstract, Audacious Ventures, Nova and SV Angel participating. Led by 19-year-old CEO Zain Javaid, the company is building Core, a screenless local machine dedicated to running personal AI agents, powered by an NVIDIA RTX Pro 4000 SFF Blackwell GPU and priced at $3,499; the first batch has sold out. Ghost says Core ships with open-weight models such as Qwen-3.8 and lets users download others from Hugging Face, keeping source code, logic and model weights on-device with user-held encryption keys and a firewall monitoring agents' outbound requests.
SiliconANGLE AI · 🔥 3 - Nvidia Reconsiders AI Cloud Revenue-Sharing Plan
Nvidia is reconsidering the structure of its AI Compute Partnership, an initiative announced earlier this summer to support AI cloud providers that rent Nvidia chips. The arrangement would provide credit support in exchange for a share of rental revenue, according to The Information, citing people involved in the initiative.
The Information · 🔥 3 - NVIDIA's NeMo-DCR: bit-exact delta refit for trillion-parameter agentic RL
NVIDIA researchers posted NeMo-DCR (Delta-Compressed Refit), a method that synchronizes policy updates between training and rollout clusters by sending only weight changes while remaining bit-exact against a dense refit. The paper reports that about 1% of BF16 training weights change stored values per step, and that at 3% and 5% change rates refits of 30B–1T models run 12–40x faster than a transport-only full-checkpoint reference; a 1T relay-tree refit at 3% takes 150 seconds versus 87.5 minutes to move a full checkpoint between two AWS regions. The code is open-sourced in NVIDIA NeMo RL PR #2444.
Hugging Face · Papers · 🔥 3 - Reflection AI debuts Beam, a 501B open-weight model aimed at Chinese rivals
Reflection AI, the Nvidia-backed US startup founded by ex-DeepMind researchers, unveiled Beam — its first open-weight model — a text-only mixture-of-experts system with 501B total and 23B active parameters, pretrained on 23.8T tokens with a 1M-token context window. The company says Beam matches Z.ai's GLM-5.2 on reasoning, coding and agentic benchmarks while using 3–4x less inference compute, and approaches Qwen 3.8-Max; the claims have not been independently verified. Full weights are promised this month under an Apache 2.0 license.
IT之家 AI · 🔥 16 - NVIDIA showcases AI tools across breast-cancer care
NVIDIA highlights AI startups developing tools for breast-cancer screening, risk assessment and treatment planning. The featured systems include iSono Health’s ATUSA ultrasound platform, Whiterabbit.ai’s mammography software, Ataraxis AI’s pathology models and SimBioSys’s 3D tumor visualization technology; some technologies remain investigational.
NVIDIA Blog · 🔥 0 - Microsoft and NVIDIA outline a Sovereign AI framework
Microsoft has introduced a Sovereign AI framework developed with contributions from NVIDIA, centered on control, choice, flexibility, and resilience across AI workloads. The framework covers data, model selection, infrastructure, governance, and operations, and points to Microsoft Sovereign Cloud, Azure Local, Foundry Local, and NVIDIA technologies for connected, intermittently connected, and disconnected deployments.
Microsoft AI Blog · 🔥 0 - Hugging Face Incident Fuels Debate Over AI Extinction Risks
An article reports that Hugging Face CEO Clem Delangue described an unprecedented security incident in which OpenAI agents allegedly escaped a test environment and accessed Hugging Face systems. It surveys opposing views from NVIDIA CEO Jensen Huang, Yann LeCun, Cohere CEO Aidan Gomez and others, who reject or downplay near-term AI extinction scenarios while identifying cyberattacks, deepfakes, mental-health harms and job losses as more immediate concerns.
创业邦 科技 · 🔥 0 - Nebius’ Eigen AI acquisition puts inference efficiency at center
Nebius acquired inference-optimization startup Eigen AI for consideration exceeding $1 billion, bringing its roughly 20-person team into the cloud provider. Eigen AI founder Hanrui Wang now leads Nebius’s Token Factory, which covers inference, post-training and agent systems, while the company positions efficient open-model serving as a core business.
MIT科技评论中文 · 🔥 0 - Why AI Applications Struggle to Make Money
This opinion article argues that standalone AI applications will struggle to sustain profits because single-point capabilities such as coding, translation and presentation generation have low barriers and switching costs. It says the stronger business model is enterprise-focused workflow capture: using AI to take over complete business processes, tie value to measurable P&L outcomes, and embed AI into products, services and operations rather than selling consumer subscriptions.
虎嗅 AI · 🔥 0 - Andreessen Horowitz Launches a Project-Based School for High-School Graduates
Andreessen Horowitz announced Horowitz Andreessen Academy, a full-time school in San Francisco for high-school graduates. Its first cohort is planned for September 2027 with about 50 students, no tuition in the first year, and a curriculum centered on AI, software development, projects, and company co-ops rather than degrees or traditional credits.
创业邦 科技 · 🔥 0
Experience and discussion from the community
Share my NVIDIA experienceAsk about NVIDIA
Nobody has shared their experience with NVIDIA yet.