NVIDIA
Start new thread
trending
•

3h ago

NVIDIA Personal AI Router - Turn your PCs and Macs into a private AI cluster

NVIDIA PAIR turns the computers on your local network into a personal AI inference cluster. Connect RTX PCs, DGX Spark systems, and Macs, then run local AI apps and agents through one endpoint while PAIR automatically routes requests to available machines. It works with Ollama and LM Studio across Windows, Linux, and macOS — while keeping prompts, files, and agent context on your network.
•

3mo ago

Nemotron 3 Ultra by NVIDIA - Powers faster, efficient reasoning for long-running agents

A 550B MoE frontier-intelligence open model built for long-running agents. It delivers 5x faster inference and lowers the cost of complex agentic tasks by up to 30% versus other open frontier models.Ultra excels at complex tasks like coding and deep research. Long-running agents spend their time planning, using tools, recovering from failures, and deciding what to do next.
•

7mo ago

NVIDIA PersonaPlex - Natural Conversational AI With Any Role and Voice

We introduce PersonaPlex, a full-duplex conversational AI model that enables natural conversations with customizable voices and roles. PersonaPlex handles interruptions and backchannels while maintaining any chosen persona, outperforming existing systems on conversational dynamics and task adherence.
•

3mo ago

NVIDIA Nemotron 3 Ultra - The first open frontier model built for agents

NVIDIA's 550B Mixture-of-Experts model with hybrid Mamba-Attention architecture, delivering 300+ tokens/sec with a 1M-token context window. Top-ranked US open-weights model on the Artificial Analysis Intelligence Index. Built specifically for multi-step agent loops where frontier reasoning at open-source economics actually matters. Available now on Hugging Face, OpenRouter, ModelScope, and build.nvidia.com as a NIM microservice.
•

6mo ago

DLSS 5 - The GPT moment for real-time computer graphics

NVIDIA DLSS 5 introduces a real-time neural rendering model that infuses game pixels with photoreal lighting and materials. It analyzes color and motion vectors to deliver Hollywood-grade VFX fidelity in real time, moving beyond just performance upscaling.
•

6mo ago

NVIDIA NemoClaw - Run autonomous agents more safely

NVIDIA NemoClaw is an open source stack that simplifies running OpenClaw always-on assistants safely. It installs the NVIDIA OpenShell runtime, part of NVIDIA Agent Toolkit, a secure environment for running autonomous agents, with inference routed through NVIDIA cloud.
•

9mo ago

CUDA 13.1 - The biggest CUDA expansion since 2006

NVIDIA CUDA 13.1 introduces the largest and most comprehensive update to the CUDA platform since it was invented two decades ago.
•

6mo ago

Nemotron 3 Super - Open hybrid Mamba-Transformer MoE for agentic reasoning

Nemotron 3 Super is NVIDIA"s open 120B model with 12B active parameters, a 1M-token context window, and a hybrid Mamba-Transformer MoE design. It is built for coding, long-context reasoning, and multi-agent workloads without the usual thinking tax.
•

2yr ago

NVIDIA Chat with RTX - Build a localized, personal AI chatbot

Chat With RTX is a demo app that lets you personalize a GPT large language model (LLM) connected to your own content—docs, notes, videos, or other data.
•

2yr ago

NVIDIA Edify 3D - Scalable High-Quality 3D Asset Generation

Using text and images as descriptions, developers and visual content creators can use NVIDIA Edify 3D to quickly generate 3D objects to create virtual worlds and prototype ideas.
12
Next