Nvidia's Vera CPU and the Olympus cores that power it: Deep dive
Nvidia unveiled its Vera CPU, a standalone AI‑centric processor featuring 88 custom ARMv9.2 cores and 1.5 TB of memory, targeting hyperscalers and cloud providers.
Coverage tone · Mixed · 0
Coverage includes a positive note on NVIDIA’s investment in GMI Cloud, a critical piece about Amazon’s Trainium chips reducing reliance on NVIDIA hardware, and a neutral mention of a self‑policing AI accord involving NVIDIA.
Nvidia unveiled its Vera CPU, a standalone AI‑centric processor featuring 88 custom ARMv9.2 cores and 1.5 TB of memory, targeting hyperscalers and cloud providers.
Nvidia-led Open Secure AI Alliance has formed a working group to develop open-source AI cybersecurity guidelines and tools.
Nvidia announced the open‑source cuFile API, which enhances GPU read/write performance for high‑speed storage systems, improving data throughput for AI and HPC workloads.
NVIDIA introduces the Rubin chip architecture, enabling AI workloads directly on personal devices and expanding its AI hardware portfolio.
The U.S. AI chip startup Etched announced a $10.3 billion valuation after securing backing from SK Hynix, positioning its custom processors to address power‑efficient inference workloads.
AMD unveiled Instella‑MoE‑16B‑A3B, a 16B parameter open‑source Mixture‑of‑Experts language model that activates only 2.8B parameters per token, trained from scratch on Instinct MI300X and MI325X GPUs.
Google Cloud’s monthly blog highlights new AI infrastructure releases, including managed Lustre, C4N VMs, and GKE Dataplane V2, all announced on July 31, 2026.
NXP Semiconductor is negotiating to buy Ambarella, a key player in autonomous‑driving camera chips, potentially reshaping the automotive AI chip market.
Safe Superintelligence, the AI lab founded by former OpenAI co‑founder Ilya Sutskever, announced a partnership with Nvidia to access its Vera Rubin GPU platform.
AI inference hardware startup Etched secured a $300 million Series C at a $10.3 billion valuation, led by Sequoia and backed by Andreessen Horowitz and others.
NVIDIA released a developer preview driver confirming the specifications of its upcoming RTX Spark N1X, a key component for its next‑generation gaming and productivity platforms.
Jeff Bezos’ Bezos Expeditions invested in British AI‑materials startup CuspAI, which partners with Nvidia to discover new semiconductor and clean‑energy materials.
Microsoft announced it will deploy AMD’s new Helios rack‑scale AI system in Azure data centers, marking a significant shift in AI infrastructure competition.
Apple’s market value surpassed Nvidia’s, reclaiming the top spot among publicly traded companies amid a surge in AI‑driven demand for its hardware and services.
NVIDIA partners with Noetra Corp. to build the Vera Rubin AI factory in Japan, delivering 140 MW of data‑center capacity and enabling open multimodal foundation models for physical AI.
Nvidia unveiled Cosmos 3 Edge, a world‑model for robots and vision AI, and announced a partnership with Japanese firms to build a physical AI ecosystem in Japan.
Former OpenAI executive Mira Murati’s startup, Thinking Machines, secured a $2 billion round led by Andreessen Horowitz, with participation from Nvidia, Cisco, and others.
NVIDIA expands its cloud gaming service by adding a new high‑performance server in Toronto, powered by the RTX 5080 GPU.
Apple has filed a lawsuit in California, claiming that OpenAI stole trade secrets related to hardware design from former Apple employees, a lawsuit that could reshape the competitive landscape for AI hardware development.
Microsoft’s Foundry platform now offers Anthropic’s Claude models on Azure, running on NVIDIA’s new GB300 Blackwell Ultra racks, delivering significant speed and memory gains for enterprise AI workloads.
The startup that gives companies a full‑stack AI agent platform has secured a $130 million Series A at a $1 billion valuation, led by Radical Ventures and backed by Nvidia Ventures, Intel Capital, and others.
Former Meta AI chief Yann LeCun has founded Advanced Machine Intelligence Labs (AMI Labs) in Paris, securing over $1 billion from Nvidia and Jeff Bezos’s wealth fund to develop a new AI architecture that moves beyond large language models and tackles real‑world physics.
Field trials demonstrate that GPU clusters can slash electricity consumption within seconds of receiving a grid signal, potentially unlocking 100 GW of capacity without building new power plants.
NVIDIA’s new two‑tower diffusion language model delivers 2.42× faster text generation than traditional autoregressive models while retaining 98.7% of baseline quality, cutting inference costs dramatically.