
A 1 million-token context window, allowing the model to retain far more information for long, multistep tasks.
Nemotron 3 Super is a high-accuracy reasoning model for multi-agent applications, while Nemotron 3 Ultra is for complex AI applications. Both are expected to be available in the first half of 2026.
NVIDIA also released today an open collection of training datasets and state-of-the-art reinforcement learning libraries. Nemotron 3 Nano fine-tuning is available on Unsloth.
Download Nemotron 3 Nano now from Hugging Face , or experiment with it through Llama.cpp and LM Studio.
DGX Spark enables local fine-tuning and brings incredible AI performance in a compact, desktop supercomputer, giving developers access to more memory than a typical PC.
Built on the NVIDIA Grace Blackwell architecture, DGX Spark delivers up to a petaflop of FP4 AI performance and includes 128GB of unified CPU-GPU memory, giving developers enough headroom to run larger models, longer context windows and more demanding training workloads locally.
Larger model sizes. Models with more than 30 billion parameters often exceed the VRAM capacity of consumer GPUs but fit comfortably within DGX Spark’s unified memory.
More advanced techniques. Full fine-tuning and reinforcement-learning-based workflows — which demand more memory and higher throughput — run significantly faster on DGX Spark.
Local control without cloud queues. Developers can run compute-heavy tasks locally instead of waiting for cloud instances or managing multiple environments.
DGX Spark’s strengths go beyond LLMs. High-resolution diffusion models, for example, often require more memory than a typical desktop can provide. With FP4 support and large unified memory, DGX Spark can generate 1,000 images in just a few seconds and sustain higher throughput for creative or multimodal pipelines.
The table below shows performance for fine-tuning the Llama family of models on DGX Spark.
As fine-tuning workflows advance, the new Nemotron 3 family of open models offer scalable reasoning and long-context performance optimized for RTX systems and DGX Spark.
Learn more about how DGX Spark enables intensive AI tasks .
Key considerations
- Investor positioning can change fast
- Volatility remains possible near catalysts
- Macro rates and liquidity can dominate flows
Reference reading
- https://blogs.nvidia.com/blog/rtx-ai-garage-fine-tuning-unsloth-dgx-spark/#content
- https://www.nvidia.com/en-us/
- https://blogs.nvidia.com/?s=
- '$100 Steam Machine' uses a cut-down PS5 APU with Bazzite — DIY console offers 60 FPS at 1080p with 16GB of GDDR6
- Now Generally Available, NVIDIA RTX PRO 5000 72GB Blackwell GPU Expands Memory Options for Desktop Agentic AI
- How to choose a CPU – A guide to picking the right processor for your PC
- Samsung eyed up for huge 8nm chip order from Intel — the Z990 chipset for Nova Lake CPUs could be Intel's 8nm debut
- Rapidus explores panel-level packaging on glass substrates for next-generation processors — aggressive plan would help it leapfrog rivals
Informational only. No financial advice. Do your own research.