Loading timeline…
Loading timeline…
Every recorded change for this company, newest first. Back to the entity page →
183 events
Events per month · 8 months shown
Nemotron 3 Nano 30B A3B: release date changed from 2025-12-14 to 2025-12-04
New model: NVIDIA-Nemotron-Nano-9B-v2 (NVIDIA)
New model: bigvgan_v2_44khz_128band_512x (NVIDIA)
New model: nvidia/Nemotron-3-Nano-Omni-30B-A3B-Reasoning-NVFP4 (NVIDIA)
New model: nemotron-3.5-asr-streaming-0.6b (NVIDIA)
New model: nvidia/NVIDIA-Nemotron-3-Super-120B-A12B-NVFP4 (NVIDIA)
New model: nvidia/Nemotron-3-Nano-Omni-30B-A3B-Reasoning-FP8 (NVIDIA)
New model: nvidia/llama-nemotron-rerank-1b-v2 (NVIDIA)
New model: NVIDIA-Nemotron-3-Super-120B-A12B-BF16 (NVIDIA)
New model: nvidia/Qwen3.5-122B-A10B-NVFP4 (NVIDIA)
New model: nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-NVFP4 (NVIDIA)
New model: nvidia/Gemma-4-31B-IT-NVFP4 (NVIDIA)
New model: nvidia/Gemma-4-26B-A4B-NVFP4 (NVIDIA)
New model: bigvgan_v2_22khz_80band_256x (NVIDIA)
New model: NVIDIA-Nemotron-3-Nano-4B-BF16 (NVIDIA)
New model: nvidia/Qwen3.6-35B-A3B-NVFP4 (NVIDIA)
New model: Nemotron 3 Nano Omni (free) (NVIDIA)
NVIDIA NIM / build.nvidia.com lists Nemotron 3 Nano 30B A3B at $0.05 in / $0.2 out per 1M tokens
NVIDIA NIM / build.nvidia.com lists Nemotron 3 Super at $0 in / $0 out per 1M tokens
NVIDIA NIM / build.nvidia.com lists Nemotron 3 Super at $0.085 in / $0.4 out per 1M tokens
NVIDIA NIM / build.nvidia.com lists Nemotron 3 Nano Omni (free) at $0 in / $0 out per 1M tokens
NVIDIA NIM / build.nvidia.com lists Nemotron 3 Ultra at $0 in / $0 out per 1M tokens
NVIDIA NIM / build.nvidia.com lists Nemotron 3 Ultra at $0.625 in / $3.125 out per 1M tokens
NVIDIA NIM / build.nvidia.com lists Nemotron 3.5 Content Safety at $0 in / $0 out per 1M tokens
NVIDIA NIM / build.nvidia.com lists Nemotron 3.5 Content Safety at $0.2 in / $0.2 out per 1M tokens
NVIDIA NIM / build.nvidia.com lists Nemotron 3.5 Lightning at $0 in / $0 out per 1M tokens
NVIDIA NIM / build.nvidia.com lists Nemotron 3.5 Lightning at $0.08 in / $0.2 out per 1M tokens
New model: NVIDIA Nemotron 3 Ultra (Preview) (NVIDIA)
New model: NVIDIA Nemotron 3.5 Lightning 30B A3B (NVIDIA)
Fireworks AI lists NVIDIA Nemotron 3 Ultra (Preview) at $0.6 in / $2.4 out per 1M tokens
Fireworks AI lists NVIDIA Nemotron 3.5 Lightning 30B A3B at $0.05 in / $0.2 out per 1M tokens
New provider: NVIDIA NIM / build.nvidia.com (NVIDIA)
NVIDIA: How Full-Stack NIM Optimizations Deliver 2.5x More Users on Nemotron 3 Ultra
NVIDIA: Skild AI Taps NVIDIA Physical AI to Teach Robots New Tasks From a Single Video
NVIDIA: Physical AI Takes the Wheel: How the World’s Robotaxi Leaders Are Building With NVIDIA Technologies
NVIDIA: High-Throughput Structure Prediction with BioNeMo Inference Runtime
NVIDIA: d-Matrix Adopts NVIDIA NVLink Fusion for Rack-Scale XPU Deployment
NVIDIA: Boots on the Ground: ‘WARDOGS’ Goes All Out on GeForce NOW at Early-Access Launch
NVIDIA: From Wafer-Out to First Token: Codifying Supply Chain Expertise with Nemotron and Palantir Foundry
NVIDIA: When to Use Encode-Prefill-Decode Disaggregation to Accelerate Multimodal Model Serving
NVIDIA: CUDA Toolkit 13.4 Adds Windows on Arm Support and Greater Control over Shared GPUs
NVIDIA: NVIDIA Brings Real-Time AI to Broadcast, Sports and Global Streaming at IBC
NVIDIA: Frontier Reasoning Reaches the Edge: How to Deploy and Optimize Models on NVIDIA Jetson
NVIDIA: How to Carry User Identity Across Federated Kubernetes and AI Platforms
NVIDIA: NVIDIA PAIR Virtual Inference Router Expands Available Compute on Your Local Network
NVIDIA: ‘NBA 2K27’ With NVIDIA DLSS 5 Leads 28 New Games Coming to GeForce NOW
NVIDIA: The Modern CUDA Toolbox in Practice: A Step-by-Step Optimization Walkthrough
NVIDIA: Co-Designing AI Models Using Speculative Decoding for Faster LLM Inference
NVIDIA: NVIDIA and CrowdStrike Strengthen Agentic Cybersecurity Frontier
NVIDIA: Building an Adaptive Agentic Cybersecurity System with NVIDIA Nemotron
NVIDIA: How to Size GPUs for AI Inference and TCO Without Overspending
NVIDIA: Run NVIDIA BioNeMo NIM Microservices for Protein Structure Prediction in Claude Science
NVIDIA: Co-Designing AI Model Attention for Fast, Interactive Long-Context Inference
Showing up to 400 events. Narrow by year or category, or use the changes feed for cursor-paged history. Times are UTC.
NVIDIA: Scale AV Perception Across Vehicle Platforms with NVIDIA Omniverse NuRec
NVIDIA: Deploy an Open Model from Checkpoint to Inference in Two Commands with NVIDIA TensorRT Model Connect
NVIDIA: Delivering Vera: NVIDIA’s First CPU Built for Agents Is Shipping Now
NVIDIA: NVIDIA NVLink Fusion Brings NVHBM to Next-Generation AI Infrastructure
NVIDIA: NVIDIA NVLink Fusion Expands With NVHBM Custom High-Bandwidth Memory
NVIDIA: How to Train a Cross-Embodiment Robot Navigation Policy with AI Agents
NVIDIA: Experiment with Qwen3.8-Flash-Next on NVIDIA GB300 NVL72 for Agentic Coding
NVIDIA: Restore LLM Inference Capacity in Seconds with Shadow Engine Recovery in NVIDIA Dynamo
NVIDIA: Leading Publishers Bring Blockbuster PC Games and Technology to NVIDIA RTX Spark
NVIDIA: CUDA Python 1.0: Stable APIs, One Foundation, Full Platform Access
NVIDIA: Giga-Scale AI and the Ethernet Evolution: How Spectrum-X Ethernet Rewrites the Rules
NVIDIA: With Groq 3 LPX in Full Production, NVIDIA Extends Vera Rubin Inference for Agents
NVIDIA: Up to 30x More Work Per Watt: NVIDIA Vera Rubin NVL72 Sets a New Efficiency Standard for AI Agents
NVIDIA: NVIDIA Vera Rubin and Blackwell Set a New Standard for Agentic AI Performance per Watt
NVIDIA: Maximizing AI Factory Performance per Watt with NVIDIA DSX MaxLPS
NVIDIA: How NVIDIA Groq 3 LPX Unlocks Ultrafast Interactivity at Long Context on NVIDIA Vera Rubin
NVIDIA: NVIDIA BlueField-4 Powers New Scale-In Network Infrastructure for Agentic AI Factories
NVIDIA: GPU-Accelerated Clustering for Financial Instruments at Scale
NVIDIA: NVIDIA AVO Reaches 100% on ARC-AGI-3, Demonstrating a Frontier-Level General-Purpose Architecture for Long-Horizon Autonomous Agents
NVIDIA: Bring the Fire: Play Games on GeForce NOW With New Firefox Browser Support
NVIDIA: Developing NVIDIA Holoscan Applications with CLI, Skills, and AI Coding Agents
NVIDIA: Evaluating AI Agent Skill Performance with NVIDIA SkillEvaluator
NVIDIA: How AI Coding Agents Can Unlock Materials Simulation with NVIDIA ALCHEMI Toolkit
NVIDIA: Run Massive-Scale UMAP in Minutes Using Multiple GPUs—Without Losing Accuracy
NVIDIA/TensorRT-LLM released v1.3.0rc23.post1: [None][ci] Recover necessary CI stages for release (#17895)
NVIDIA/TensorRT-LLM released v1.3.0rc22.post1: [None][chore] Bump version to 1.3.0rc22.post1
NVIDIA: Developing Nemotron 3.5 Lightning NVFP4 with QAD Using NVIDIA Model Optimizer
NVIDIA: Serve Qwen3.8-2.4T-A95B, a 2.4T-Parameter Model, with Configurable Reasoning on NVIDIA GB300 NVL72
NVIDIA: How to Choose Full-Stack Observability for NVIDIA AI Factories
NVIDIA: NVIDIA JetPack 7.2.1 Adds Agentic Video Skills and T3000 Emulation
NVIDIA: NVIDIA Nemotron 3.5 Lightning Delivers Fast, Accurate Specialized Task Execution for Long-Running Agents
NVIDIA: Route AI Agent Workloads Across Models with NVIDIA NeMo Switchyard
NVIDIA: Run Local Agentic AI Workflows with Meta’s Muse Glimmer on NVIDIA
NVIDIA: Beyond VLAs: How World Action Models Reshape Robot Manipulation
NVIDIA: Generate Trajectories, Reasoning Traces, and Auto-Labels with NVIDIA Alpamayo 2 Super
NVIDIA: NVIDIA Vera Storage Benchmarks: Faster Encryption, Compression, Integrity Checking, and Recovery for AI-Native Storage
NVIDIA: How to Run Isolated Tenant Kubernetes Clusters on Shared GPU Infrastructure
NVIDIA: NVIDIA Video Codec SDK 13.1: Zero-Copy Transcode, AV1 B-Frames, and Frame-Accurate Seek
NVIDIA: Run High-Performance Core Math at Scale with NVIDIA nvmath-python
NVIDIA: NVIDIA Exemplar Cloud: Lessons for Unlocking Full Performance on AI Infrastructure
NVIDIA: How to Self-Host a Validated AI Coding Assistant with NVIDIA NeMo Guardrails
NVIDIA: Developing Healthcare Robotics with GPU-Native Medical Physics Simulation
NVIDIA: NVIDIA Ising Enables Fully Automated Quantum Computer Calibration with Enhanced In-Context Learning
NVIDIA: NVIDIA Nemotron 3 Ultra Leads Open Models on Accuracy and Efficiency in Agentic RTL Coding
NVIDIA: Advancing Semiconductor Innovation Across Materials Engineering and Manufacturing
NVIDIA: ModelExpress: Distributing Model Artifacts at the Speed of Light
NVIDIA: Debugging Ray Tracing Applications Using NVIDIA OptiX Toolkit
NVIDIA: Start Customizing NVIDIA Nemotron 3 Nano with Prime Intellect Lab in Minutes
NVIDIA: Make Long-Running NVIDIA TensorRT Engine Builds Observable and Cancelable in Python or C++
NVIDIA: Setting a World Record for MoE Pre-Training on NVIDIA GB300 NVL72
NVIDIA: Inside NVIDIA Rubin GPU Architecture: Powering the Era of Agentic AI
NVIDIA: NVIDIA Vera CPU: Olympus Cores Built for Maximum Single-Thread Performance in Agentic AI
NVIDIA: Integrate NVIDIA Omniverse RTX Sensor Simulation Into Existing Apps
NVIDIA: Q&A: How Capcom Brought Path Tracing to RE ENGINE Across PRAGMATA and Resident Evil Requiem
NVIDIA: Integrating Context-Aware Video AI Agents Into Enterprise Workflows
NVIDIA: Scaling Agentic AI Factories Through Extreme Co-Design with NVIDIA BlueField
NVIDIA: Build a Multi-Camera 3D Tracking Application with NVIDIA DeepStream 9.1 Skills
NVIDIA: Building Faster Cryptography with Carryless Multiplication in NVIDIA CUDA 13.3
NVIDIA: Lessons From the Leaderboard: What 5,000+ Kagglers Taught Us About Improving AI Reasoning
NVIDIA: How to Run an Autoresearch Workflow with RL Agent Skills and NVIDIA NeMo
NVIDIA: NVIDIA Ising Decoding Cuts Color Code Logical Error Rates by Over 300x
NVIDIA: How to Evaluate General-Purpose Robot Policies for Real-World Deployment
NVIDIA: Reducing High-Bandwidth Memory Bottlenecks in JAX-Based LLM Training with Host Offloading
NVIDIA: Kernel Fusion in NVIDIA CUDA: Optimizing Memory Traffic and Launch Overhead
NVIDIA: Accelerating End-to-End Co-Folding Performance with NVIDIA BioNeMo Agent Toolkit
NVIDIA: Synthetic Data Generation for Financial AI Research with NVIDIA NeMo
NVIDIA: A Practical Guide to GPU-Initiated Communication for Molecular Dynamics at Scale
NVIDIA: Create a LangChain Deep Agents Harness Profile for NVIDIA Nemotron 3 Ultra to Improve Performance
NVIDIA: Running Low-Latency Analytical Workloads with GPU-Accelerated Presto on NVIDIA GB200 NVL72
NVIDIA: NVIDIA Vera CPU Boosts AI Factory Throughput to Accelerate Agentic Workloads
NVIDIA: Develop Humanoid Robot Policies End-to-End with NVIDIA Isaac GR00T
NVIDIA: Building an Analysis AI Agent for Industrial Alarm Management with NVIDIA Nemotron
NVIDIA: Maximize Spectral Efficiency with AI-Native RAN and NVIDIA AI Aerial
NVIDIA: Enhancing Goodput in Large-Scale LLM Training with Nonuniform Tensor Parallelism
NVIDIA: Mastering Agentic Techniques: AI Agent Reinforcement Learning
NVIDIA: Optimizing a Neural Reconstruction Pipeline Using NVIDIA Nsight Developer Tools
NVIDIA: Deploy a Production-Ready NVIDIA AI-Q Blueprint on Oracle Cloud Infrastructure
New model: Nemotron 3.5 Content Safety (NVIDIA)