NVIDIA is a pioneer of GPU computing, dominating the hardware market for AI acceleration, training, and inference with its high-performance Hopper and Blackwell architectures.
Directly governs the hardware efficiency and hardware-level token throughput when deploying ai hardware computing clusters, deep learning acceleration, and gaming graphics cards; optimizing NVIDIA is a major factor in compute cost budgeting.
NVIDIA is a leading global technology company that pioneered GPU-accelerated computing. Initially known for graphics cards, NVIDIA became the primary hardware provider for the modern artificial intelligence revolution. Its compute platform (including CUDA software and H100/A100 and Blackwell GPU architectures) serves as the computational backbone for training and serving frontier LLMs.
NVIDIA provides both the high-performance GPU hardware and the CUDA software ecosystem, making it the industry standard for developing deep learning systems.
Hopper (e.g. H100) is NVIDIA's previous-generation AI enterprise GPU. Blackwell is its successor, offering 30x faster LLM inference performance.
Reference this definition in your articles, research, or documentation to credit this source:
I'm excited to announce that NVIDIA has agreed to acquire Hugging Face for $12,930,300,000. Together, we will scale Hugging Face's platform, strengthen its...
Agentic artificial intelligence security startup Capsule Security Ltd. today released a detection system built on two Nvidia Corp. Nemotron models it fine-tuned itself, in what it calls an "AI circuit breaker" for rogue AI agent. The models judge an agent's intended action in the moment before...
When enterprise buyers build out their next AI accelerator evaluation list this cycle, they're more likely to put a non-Nvidia chip on it than Nvidia's own...
Follow this step-by-step guide to deploy Kimi K3 on CoreWeave Dedicated Inference on NVIDIA GB300 NVL72.
Artificial intelligence hardware startup Starcloud Inc. today announced that it has raised $250 million in funding at a $2.3 billion valuation. Investment firm Manhattan West led the deal. It was joined by more than a dozen other backers, including Nvidia Corp. and Cisco Investments. The cash...
Nvidia research shows that AI agent can perform well, and not go off the deep end, through fine-tuning, even if the AI model isn't that great at the task.
When an agentic AI system hands a task from a small model to a larger one - or back down again - it pays a steep tax: the receiving model has to recompute...
Artificial intelligence startup Groq Inc. today announced that it has raised $350 million in additional funding. The Series A round was led by returning backer Disruptive. Grok stated that Nvidia Corp. plans to join the round later down the line, but didn't specify how much the chip giant will...
NVIDIA Nemotron 3.5 Lightning, an open model built for high-volume agentic workloads, is now available in Amazon SageMaker JumpStart. This post shows how to...
Groq raised $350 million at a $3.5 billion valuation as the former AI chipmaker pivot to a neocloud business and expands its Nvidia-powered data center...
Nvidia Corp. is no longer just selling technology. It is helping create a financial asset class around artificial intelligence compute. In our last Breaking Analysis, we argued that AI can be technologically transformative and still produce a capital bubble. Our thesis was simply that the bubble...
Indonesia is taking charge of its AI future. This week, the Ministry of Communication and Digital Affairs (Komdigi), Indosat Ooredoo Hutchison (Indosat or...
Nvidia has a plan to make sure its GPU won't lose value. It wants to convince a new crop of financiers to keep lending for AI buildouts.
Production AI factory lifecycle series, Part two: how CoreWeave operates the stack at scale, from goodput to reliability, and carries it into NVIDIA Vera Rubin NVL72 readiness.
A production AI factory is a lifecycle, not a handoff. Part one: how NVIDIA and CoreWeave co-design an integrated and optimized tech stack and validate it before customers deploy.
NVIDIA founder and CEO Jensen Huang is ranked No. 1 on Glassdoor's Best CEOs list for 2026. In the just-released ranking, recognition is earned directly from...
We announced partnerships with Apollo, BlackRock, Blackstone, Brookfield, Goldman Sachs and KKR to establish independent financing platforms designed to...
IBM Corp. will provide infrastructure to artificial intelligence startup Together AI Inc. as part of a $240 million deal announced today. The partnership comes a few weeks after the latter company raised $800 million in funding. The consortium that provided the capital included Nvidia Corp...
River AI Inc., a startup that helps enterprises customize open-source artificial intelligence models, has raised $1.1 billion in early-stage funding. The company stated in today's announcement that it received the capital over two rounds, a seed and a Series A. General Catalyst and AMP PBC were...
The open source ecosystem is making it easier for AI enthusiasts and developers to build, customize and run increasingly capable agents locally. Throughout...
As AI shifts from chatbot to autonomous agent, open models are serving market demands for full control over where AI runs and how it's deployed and evolves...
Artificial intelligence silicon and software giant Nvidia Corp. today announced two new services: a highly customizable Nemotron model and an agentic AI model router named NeMo Switchyard. As enterprises find themselves drowning in artificial intelligence model options, the question is no longer...
Some of Wall Street's biggest financial firms are partnering with Nvidia Corp. to pour a half-trillion dollars of funding into the artificial intelligence industry's massive infrastructure buildout. Nvidia said today it has struck deals with Apollo Global Management Inc., BlackRock Inc...
UC Berkeley AI Research demonstrates K-Search automated kernel transpilation from NVIDIA CUDA to Apple MLX hardware primitives.
NVIDIA announces new Jetson Thor edge modules for autonomous robotics and real-time physical AI inference.
Projects that want to share a GPU on Kubernetes have to work around an API instead of with it. The device plugin interface could count devices, and that was the whole vocabulary: nvidia.com/gpu: 1. It meant one...
Anthropic and OpenAI are racing to scale up while reducing dependence on Nvidia.
In July, NVIDIA joined more than 200 companies and organizations in signing "Open Weights and American AI Leadership," an open letter arguing that AI...
NVIDIA is participating in the U.S. National Science Foundation's (NSF) State and Regional Artificial Intelligence Infrastructure Hubs program, an effort...
For robotaxis and other autonomous vehicles (AVs), the hardest problems aren't the everyday scenarios. They're the rare, complex situations that are...
As artificial intelligence applications become ever hungrier for faster access to data, Nvidia Corp. today announced it is open-sourcing the application programming interface for its powerful cuFile vertical data storage stack, enabling millisecond data access. The company also announced a...
Valar Atomics raised $1 billion at a $6 billion valuation after signing a development deal with Nvidia in June.
Docker joins NVIDIA's Open Secure AI Alliance to help build the security, governance, and trust frameworks that agentic AI systems demand.
This week on Uncanny Valley, we discuss the open- vs. closed-source debate in AI, key players in White House AI policy, and how to stop your chatbot logs...
CoreWeave deploys liquid cooled switching for NVIDIA Vera Rubin NVL72 delivering 100% higher switching performance at 1.64 Pb/s per rack compared to air cooled switches.
As the United States government debates new artificial intelligence rules and regulations, Anthropic PBC and Nvidia Corp. are drawing a line against blanket bans on open-weight models, urging regulators to focus instead on specific risks and misuse. An open-weight model is one where the developer...
Coalition of 37 tech giants develops open-source defenses and sandboxing frameworks following Hugging Face repository investigation.
Nvidia Corp. is reportedly in talks to backstop a $250 billion loan for OpenAI Group PBC. The Wall Street Journal on Sunday cited sources as saying that the financing would cover the cost of a new data center campus in Ohio. According to CNBC, the site is expected to provide 10 gigabytes of...
The artificial intelligence competition is no longer between large language models, but rather between AI systems. Advanced Micro Devices Inc. has shifted from being primarily a chip business to a secondary player in the race to build a system of intelligence. At the company's annual event last...
Nvidia Corp. today announced the launch of the Open Secure AI Alliance, a new organization founded by technology, cloud computing and cybersecurity leaders to build and share open artificial intelligence tools. As the capability and strength of AI tools grows and shapes the technology industry...
After two years in stealth, Safe Superintelligence has announced a long-term partnership with Nvidia as it prepares to scale to its next phase.
The complexity of modern chip design continues to grow as engineering teams work to develop increasingly sophisticated CPUs, GPU and AI systems. To help...
AI hardware competition entered a sharper phase this week as Advanced Micro Devices Inc. used its flagship AI event to argue it isn't merely chasing Nvidia Corp. - it intends to lead the market outright. The shift marks a departure from years of AMD positioning itself as the pragmatic second...
A group of tech firms has released an open letter that calls on policymakers not to ban open-source artificial intelligence models. The development follows a report that some Trump administration officials sought to limit the use of such algorithm. Many of the world's most popular open-source AI...
AI companies, including Nvidia and Mistral, urge policymakers to avoid broad restrictions on open-weight AI model as Washington debates responses to Chinese...
Maybe Nvidia ultimately won't win the entire AI enchilada, even as it reminded everyone this week that it offers everything you need to build AI factories. This week, Advanced Micro Devices also announced a raft of new graphics processing units, central processing units and other processors for...
At this week's AI Summit in San Francisco, South Korean President Jae Myung Lee and some of the country's top business leaders and researchers are meeting...
AMD is challenging its chipmaker rival with a new rack-scale system that will start shipping to customers later this year.
Advanced Micro Devices Inc. is pushing harder than ever to grab even more market share from Nvidia Corp. in the artificial intelligence chip industry. At its Advancing AI 2026 event today in San Francisco, the chipmaker announced a slate of updated hardware, including its next-generation AMD...
CoreWeave publishes first-ever measured silicon performance of NVIDIA Vera Rubin NVL72, marking a new milestone in AI infrastructure.
The race to build AI infrastructure systems has moved beyond chip specifications into a battle over entire rack-scale platforms, as inference and agentic workloads redefine what counts as a computer. That shift is forcing challengers, once judged purely on GPU benchmarks, to prove they can ship...
On the surface, this week's Vera Rubin launch is another major platform moment for Nvidia Corp., as the company maintains a steady drumbeat of artificial intelligence infrastructure innovation. Nvidia is positioning Vera Rubin as a full-stack system designed to improve performance per watt and...
The AI era runs on AI infrastructure. Many of these advanced systems are built and tested in Texas. Wistron opened its first U.S. manufacturing facility...
NVIDIA Vera Rubin is here, and it's going gigascale. Vera Rubin NVL72 production is ramping up with racks running at partners CoreWeave, Google Cloud...
AI has entered the gigascale era. The world's most advanced AI factories are bringing together hundreds of thousands of GPU and CPUs to train frontier...
Artificial intelligence chip king Nvidia Corp. today revealed a fresh trove of performance benchmarks and architectural milestones for its next-generation Vera Rubin platform as it edges closer to global availability. The new numbers are impressive, but on a higher level they also underscore the...
CuspAI Ltd., a startup using artificial intelligence to discover new materials, today announced that it has raised $450 million in funding. The U.K.-based company also formed a chemical research consortium that includes Nvidia Corp. and Samsung Electronics Co. CuspAI's Series B round values it at...
In this post, we show how Amazon Quick can serve as the business-user front door for specialized agent workflows. We use the NVIDIA NeMo Agent Toolkit to...
Artificial intelligence companies are racing to control the infrastructure, data and software layers that will power enterprise intelligence. Nvidia Corp. remains far ahead in accelerated computing, but Advanced Micro Devices Inc., Broadcom Inc. and other challengers are positioning themselves...
From open models to real-time simulation, AI and graphics breakthroughs are transforming media, content creation and robotics.
Erin Davis calls it the "SuperDuperPOD." That's two things in one name: pharmaceutical giant Bristol Myers Squibb (BMS) already runs one of the largest AI...
The Compute Exchange Inc., a procurement marketplace for reserved graphics processing unit capacity, today launched a dedicated marketplace for used and refurbished GPU, extending the platform into physical artificial intelligence hardware sourcing. The service connects buyers with suppliers of...
Lowest cost per token from extreme codesign maximizes intelligence per dollar for post-training in the agentic era.
Nvidia Corp. has unveiled Cosmos 3 Edge, a compact world model built to run vision reasoning and robot control directly on edge devices, alongside a wave of partnerships that pushes its physical artificial intelligence platform deeper into Japan's robotics and manufacturing base. The model can...
In a special editorial discussion hosted by Dave Vellante and Bob Laliberte, Nvidia Corp. networking chief Gilad Shainer explains why agentic inference turns the network into part of the computer. We believe Nvidia is materially ahead of the field, but in this Special Breaking Analysis we...
As enterprises move artificial intelligence into production, AI factory networking is becoming a core part of the infrastructure equation, shaping performance, scalability and cost. That shift is the focus of a new analysis by Bob Laliberte, principal analyst at theCUBE Research. The analysis...
General-purpose robots and autonomous machines are moving from research labs to real-world mass-market deployment, creating demand for compact...
Home to leading manufacturers, robotics pioneers, infrastructure builders and iconic gaming companies, of course, Japan is one of the world's centers of AI...
In this post, we explore what makes the Nemotron 3 architecture unique, walk through the fine-tuning techniques available, and show you step-by-step how to...
The company is using the cash to open an office in the Bay Area and compete for talent there, "strengthening its position at the heart of the world's leading...
Artificial intelligence training startup Prime Intellect Inc. has raised $130 million in funding from a group of prominent investors. The consortium included Nvidia Corp.'s NVentures, Intel Capital and Dell Technologies Capital. They were joined by more than a dozen others, including Cloudflare...
NVIDIA Nemotron 3 Ultra is offering leading performance at lower cost than top closed models with the largest and most widely adopted AI agent orchestration...
Max single-threaded CPUs at scale are a new category of CPUs built for the agentic AI era. Across the creation and deployment of an agentic system, the CPU...
Open source AI has shown how quickly developers can innovate when models, data and tools are shared. Robotics has the same opportunity, but advancements in...
Physical artificial intelligence is becoming an industrial robotics problem. The market is shifting from software-only automation toward machines that must sense, decide and act in physical settings. That raises the bar for safety, economics and reliability, especially as companies such as Nvidia...
As AI moves from model development to production inference, compute demand is accelerating and shifting toward continuously operating AI factories that...
We're excited to introduce US-based frontier open-weight models in AWS GovCloud (US). With this release, Amazon Bedrock now supports OpenAI's open-weight GPT...
Dynamic Resource Allocation (DRA) recently reached GA in Kubernetes v1.35, and I believe many of us are eager to give it a try. Adding to the momentum, NVIDIA has moved dra-driver-nvidia-gpu into Kubernetes SIGs, with the...
Life sciences has entered an era of computational scale, and for more than a decade, NVIDIA has built the full GPU-accelerated computing stack - spanning...
As organizations move from AI pilots to production AI factories, infrastructure decisions have shifted from peak chip specifications to cost per token: how...
Apple. Anthropic. Disney Research. Google. Meta. Microsoft. NVIDIA. OpenAI. Few places outside Silicon Valley can claim R&D hubs from all of these companies...
Showcasing the importance of open source innovation in American AI, Palantir's new intelligent engine - introduced today - uses NVIDIA Nemotron open models to...
Nvidia has dominated the AI chip market for years, but the era of total dependence might be ending. OpenAI just shared its plans to spice things up with...
Nvidia has dominated the AI chip market for years, but the era of total dependence might be ending. OpenAI just shared its plans to spice things up with...
This post shows you how to configure training jobs on Amazon SageMaker AI to get the most out of Blackwell's architecture on AWS. You learn how to select...
Building AI systems at scale is demanding, requiring low-latency inference, fast vector search, strong GPU price-performance and infrastructure that can grow...
Telecom operators have seen remarkable returns from using generative AI to automate network management, customer care and back-office operations. Most of that...
For the past two years, the U.S. National Science Foundation's National Artificial Intelligence Research Resource (NAIRR) pilot program has driven innovative...
Flexion Robotics, a startup founded by ex-Nvidia engineers, has a clever way of training robots to do useful work.
Mission, Vision and Veritas - new Los Alamos National Laboratory (LANL) supercomputers to be built with HPE and NVIDIA - are tapping NVIDIA Vera CPUs to...
At the ISC conference running in Hamburg this week, NVIDIA is introducing new software that speeds AI for science, from chemistry and materials discovery to...
The next era of AI will not be defined by compute alone. Its growth will be determined by energy. As accelerated computing scales across AI factories, agentic...
Hot tubs sit at about 38 to 40 degrees Celsius, warm enough that most people can only soak for about 15 minutes. NVIDIA's newest AI servers can run their...
The digital era gave the advertising and marketing industry speed; the AI era is giving it autonomous operations. For companies building next-generation...
Palantir Technologies Inc. Chief Executive Alex Karp appeared to go into meltdown mode during an interview with CNBC today where for 20-odd minutes he went off script after being asked to discuss his company's ongoing deal with chip maker Nvidia Corp. The conversation turned to the AI industry...
A year ago at NVIDIA GTC Paris at VivaTech, France laid out plans to advance local AI - from new AI factories and national compute capacity to open frontier...
Enterprises are moving agentic AI from proof of concept to production - and the next generation of AI factories are built for the era of agents. At HPE...
Lambda's GB300 NVL72 Llama 3.1 8B MLPerf Training v6.0 submission improved performance by 18.7% over Lambda's previous result, achieving the fastest convergence on this round's workload on GB300 NVL72. In addition, Lambda achieved the fastest result among single-node HGX B200 submissions for...
AgentPerf from Artificial Analysis, the industry's first agentic AI benchmark, gives developers, enterprises and infrastructure providers a clear way to...
Today, Google DeepMind released DiffusionGemma - an experimental open model built for exceptionally fast text generation. NVIDIA has optimized DiffusionGemma...
NVIDIA GPU with Confidential Computing are now used for confidential inference in Apple's Private Cloud Compute (PCC), as it expands beyond Apple's data...