# NVIDIA NeMo Projects at AI Tinkerers

> Canonical HTML: https://aitinkerers.org/technologies/nvidia-nemo
> Markdown URL: https://aitinkerers.org/technologies/nvidia-nemo.md
> Technology record last updated: 2026-03-05T13:07:02Z
> Generated: 2026-09-22T07:51:39Z

Modular software suite for building, customizing, and deploying enterprise-scale generative AI agents and models.

NVIDIA NeMo is the comprehensive, modular software suite for managing the entire AI agent lifecycle. It provides microservices and a cloud-native framework for data curation, model customization, and secure deployment across any GPU-accelerated infrastructure. Key components like NeMo Curator, NeMo Customizer, and NeMo Evaluator streamline the workflow: you process high-quality data, fine-tune models like LLMs and VLMs using techniques like p-tuning, and benchmark performance. NeMo supports multi-node and multi-GPU scaling, ensuring production-ready performance for conversational AI, ASR, NLP, and multimodal applications. This is the toolkit for delivering continuously optimized, secure AI at enterprise scale.

- Official technology site: https://www.nvidia.com/en-us/ai-data-science/generative-ai/nemo/
- Public AI Tinkerers demos and talks: 3
- Result page: 1 of 1

## Recent Public Talks and Demos

### [Silent Notetaker: no backend, no account, no upload](https://columbus.aitinkerers.org/talks/rsvp_2_Cj7Fhun3Y)

Silent Notetaker is a meeting notetaker that runs entirely in the browser: it transcribes the conversation live, pulls out decisions, action items, and open questions as they happen, with no audio ever leaving the machine. The whole app is a single HTML file, no backend and nothing to sign into. I'll demo it live with a real mic: transcription and speaker labels appearing in real time, notes self-categorizing, and an on-device LLM suggesting the next question to ask. Then review architecture (how the speech model, the speaker model, and the question model share one machine without fighting each other), show the single-file source for people to try out, modify and make it their own.

- Event context: AI Tinkerers - Columbus June Meetup — 2026-06-01 — Columbus
- Public talk page: https://columbus.aitinkerers.org/talks/rsvp_2_Cj7Fhun3Y

### [Personal AI Supercomputers: From Cloud Dependency to Local AI](https://paris.aitinkerers.org/talks/rsvp_k6HVVqqlz3s)

AI development is entering a new era where developers no longer need massive cloud clusters to build and run advanced models. A new class of AI-native personal supercomputers, powered by NVIDIA’s Grace-Blackwell architecture, is bringing datacenter-grade AI capabilities directly to the developer desk. In this talk, we will explore how systems like NVIDIA DGX Spark and Lenovo ThinkStation PGX are reshaping the way AI engineers prototype, train, and deploy models locally — from large language models to multimodal and agentic AI systems. We will explain the architecture behind the GB10 Grace-Blackwell superchip, unified memory systems, and the NVIDIA AI software stack that makes these platforms powerful tools for experimentation and enterprise AI development. This session will answer key questions such as: Why are AI personal supercomputers emerging as a new category of computing? What problems do developers face today with cloud-only AI development? How do DGX Spark and Lenovo ThinkStation PGX enable developers to run models up to hundreds of billions of parameters locally? How does unified memory and low-precision computing (FP4/FP8) accelerate modern AI workloads? What role does the NVIDIA AI ecosystem (NeMo, NIM, Blueprints, CUDA libraries) play in building AI agents and applications? How does the AI development workflow evolve from prototyping to deployment across personal, enterprise, and cloud systems? We will also demonstrate how these systems enable developers to move seamlessly from experimentation to production-scale AI while maintaining performance, security, and cost efficiency.

- Event context: High-Performance Local AI Development: Kick-off ThinkStation PGX — 2026-03-17 — Paris
- Public talk page: https://paris.aitinkerers.org/talks/rsvp_k6HVVqqlz3s

### [Parakeet's French Accent: No GPU Required to Sound Fancy](https://montreal.aitinkerers.org/talks/rsvp_ElMZDjHOr7I)

How the model react with something else than English and without GPU

- Event context: AI Tinkerers Montreal: Demo Night — November 20, 2025 — 2025-11-20 — Montreal
- Public talk page: https://montreal.aitinkerers.org/talks/rsvp_ElMZDjHOr7I

## Related Technologies

- [Browser](https://aitinkerers.org/technologies/browser) ([Markdown](https://aitinkerers.org/technologies/browser.md)) — 5 public demos
- [CUDA](https://aitinkerers.org/technologies/cuda) ([Markdown](https://aitinkerers.org/technologies/cuda.md)) — 15 public demos
- [DGX Spark](https://aitinkerers.org/technologies/dgx-spark) ([Markdown](https://aitinkerers.org/technologies/dgx-spark.md)) — 1 public demo
- [Grace-Blackwell](https://aitinkerers.org/technologies/grace-blackwell) ([Markdown](https://aitinkerers.org/technologies/grace-blackwell.md)) — 2 public demos
- [IndexedDB](https://aitinkerers.org/technologies/indexeddb) ([Markdown](https://aitinkerers.org/technologies/indexeddb.md)) — 1 public demo
- [Mistral](https://aitinkerers.org/technologies/mistral) ([Markdown](https://aitinkerers.org/technologies/mistral.md)) — 24 public demos
- [NGC](https://aitinkerers.org/technologies/ngc) ([Markdown](https://aitinkerers.org/technologies/ngc.md)) — 1 public demo
- [onnxruntime-web](https://aitinkerers.org/technologies/onnxruntime-web) ([Markdown](https://aitinkerers.org/technologies/onnxruntime-web.md)) — 1 public demo
- [Python](https://aitinkerers.org/technologies/python) ([Markdown](https://aitinkerers.org/technologies/python.md)) — 662 public demos
- [Speech-to-Text](https://aitinkerers.org/technologies/speech-to-text) ([Markdown](https://aitinkerers.org/technologies/speech-to-text.md)) — 10 public demos
- [Streaming](https://aitinkerers.org/technologies/streaming) ([Markdown](https://aitinkerers.org/technologies/streaming.md)) — 6 public demos
- [Transformers](https://aitinkerers.org/technologies/transformers) ([Markdown](https://aitinkerers.org/technologies/transformers.md)) — 148 public demos
- [WebAssembly](https://aitinkerers.org/technologies/webassembly) ([Markdown](https://aitinkerers.org/technologies/webassembly.md)) — 11 public demos
- [WebGPU](https://aitinkerers.org/technologies/webgpu) ([Markdown](https://aitinkerers.org/technologies/webgpu.md)) — 8 public demos
