Talks & Publications
Where I've written, spoken, taught, and rambled about AI.
Book
Research
Nemotron 3 Ultra: Open, Efficient Mixture-of-Experts Hybrid Mamba-Transformer Model for Agentic Reasoning ↗
Jun 2026arXiv:2606.15007
550B-parameter MoE hybrid Mamba-Transformer with 1M-token context, built for high-throughput agentic reasoning.
Nemotron 3 Super: Open, Efficient Mixture-of-Experts Hybrid Mamba-Transformer Model for Agentic Reasoning ↗
Apr 2026arXiv:2604.12374
120B (12B active) LatentMoE hybrid Mamba-Attention model — first in the family pre-trained in NVFP4, with MTP layers for native speculative decoding.
NVIDIA Nemotron 3: Efficient and Open Intelligence ↗
Dec 2025arXiv:2512.20856
Technical report for the Nemotron 3 family — MoE hybrid Mamba-Transformer models (Nano, Super, Ultra) with NVFP4 training, LatentMoE, and 1M-token context.
Nemotron 3 Nano: Open, Efficient Mixture-of-Experts Hybrid Mamba-Transformer Model for Agentic Reasoning ↗
Dec 2025arXiv:2512.20848
NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model ↗
Aug 2025arXiv:2508.14444
Llama-Nemotron: Efficient Reasoning Models ↗
May 2025arXiv:2505.00949
Articles
NVIDIA Nemotron 3 Ultra Powers Faster, More Efficient Reasoning for Long-Running Agents ↗
Jun 2026NVIDIA Technical Blog
Building NVIDIA Nemotron 3 Agents for Reasoning, Multimodal RAG, Voice, and Safety ↗
Mar 2026NVIDIA Technical Blog
Introducing Nemotron 3 Super: An Open Hybrid Mamba-Transformer MoE for Agentic Reasoning ↗
Mar 2026NVIDIA Technical Blog
How to Train an AI Agent for Command-Line Tasks with Synthetic Data and Reinforcement Learning ↗
Jan 2026NVIDIA Technical Blog
How to Build a Voice Agent with RAG and Safety Guardrails ↗
Jan 2026NVIDIA Technical Blog
Inside NVIDIA Nemotron 3: Techniques, Tools, and Data That Make It Efficient and Accurate ↗
Dec 2025NVIDIA Technical Blog
Develop Specialized AI Agents with New NVIDIA Nemotron Vision, RAG, and Guardrail Models ↗
Oct 2025NVIDIA Technical Blog
Build More Accurate and Efficient AI Agents with the New NVIDIA Llama Nemotron Super v1.5 ↗
Jul 2025NVIDIA Technical Blog
Build an AI Agent with Expert Reasoning Capabilities Using the DeepSeek-R1 NIM ↗
Feb 2025NVIDIA Technical Blog
Mastering LLM Techniques: Evaluation ↗
Jan 2025NVIDIA Technical Blog
Deploying Fine-Tuned AI Models with NVIDIA NIM ↗
Nov 2024NVIDIA Technical Blog
All NVIDIA Technical Blog posts → ↗
2023–NVIDIA Developer Blog
Talks
Agentic AI Summit × Generative AI Summit — Speaker ↗
Nov 2026AI Accelerator Institute · Toronto
Upcoming — co-located summits for engineers building industry-ready agentic and generative AI.
Agentic AI Summit (Virtual) — Speaker ↗
Jul 2025ODSC
RAG 101: Building an Open-Source ChatGPT for Your Data — Workshop ↗
2024ODSC East
Podcasts
Videos & Courses
Nemotron Labs Livestream ↗
WeeklyNVIDIA AI · YouTube
Weekly livestream I host on the NVIDIA AI channel, 11 AM PT — ask-the-experts Q&As and deep dives with the people building Nemotron.
The AI Engineering Bootcamp ↗
2022–2026AI Makerspace
Cohort-based bootcamp I co-created and taught, covering prompt engineering, RAG, fine-tuning, agents, and evals — build, ship, share. Now a book (Wiley, 2026).
Chris Alexiuk on YouTube ↗
2022–YouTube
Videos on LLMs, fine-tuning (LoRA and friends), and ML techniques.
AI Makerspace on YouTube ↗
2023–2026YouTube
A deep back catalog of live events on AI engineering I co-hosted — LLMs, agents, evals, and the latest tooling.