Join our Discord Server

NVIDIA

Deploying LLM Inference on Kubernetes with the NVIDIA GPU Operator: A Step-by-Step Tutorial

Serving large language models efficiently is one of the dominant themes at recent KubeCon and CloudNativeCon events, and the NVIDIA GPU...
Collabnix Team
1 min read

Top 5 Reasons to Buy NVIDIA DGX Spark

The AI computing landscape is evolving at an unprecedented pace, and NVIDIA has once again redefined what’s possible at the desktop...
Collabnix Team
4 min read
How to Run GPU Workloads on Kubernetes with NVIDIA: A Step-by-Step Guide

How to Run GPU Workloads on Kubernetes with NVIDIA: A Step-by-Step Guide

Learn how to leverage NVIDIA GPUs to run high-performance workloads on Kubernetes clusters.
Collabnix Team
7 min read

Building AI Agents on DGX Spark with Kubernetes: A Complete Tutorial

The NVIDIA DGX Spark is a compact, personal AI supercomputer powered by the GB10 Grace Blackwell Superchip. With 128GB of unified...
Collabnix Team
6 min read

PersonaPlex-7B-v1 NVIDIA: Human-Like Speech Model

Discover PersonaPlex-7B-v1 NVIDIA Speech Model If you’ve ever used a voice assistant and felt that awkward pause — where you speak,...
Collabnix Team
4 min read

Top 5 Reasons Why I’m Attending NVIDIA GTC 2026 (And Why You Should Too!)

Top Reasons to Attend NVIDIA GTC 2026 If there’s one AI event you absolutely cannot afford to miss this year, it’s...
Ajeet Raina
5 min read

NVIDIA Jetson AGX Thor: Top 5 Use Cases with Tutorials

The NVIDIA Jetson Thor represents a paradigm shift in edge computing for physical AI. Powered by the 2560 Blackwell-based CUDA cores,...
Collabnix Team
16 min read

Exploring the Llama 4 Herd and what problem does it solve?

Hold onto your hats, folks, because the world of Artificial Intelligence has just been given a significant shake-up. Meta has unveiled...
Adesoji Alu
14 min read

Getting Started with NVIDIA Dynamo: A Powerful Framework for Distributed LLM Inference

In the rapidly evolving landscape of generative AI, efficiently serving large language models (LLMs) at scale remains a significant challenge. Enter...
Collabnix Team
3 min read

Getting Started with NVIDIA Jetson Orin Nano Super – Generative AI Supercomputer

NVIDIA has just reinvented edge computing with its latest offering – the Jetson Orin Nano Super Developer Kit. This isn’t just an...
Ajeet Raina
5 min read

How to Run DeepSeek-V3 Locally on Ubuntu with Python 3.11: A Step-by-Step Guide

Quantizing DeepSeek-V3 for Smaller GPUs Large language models (LLMs) like DeepSeek-V3 offer incredible capabilities, but their size often makes them challenging...
Adesoji Alu
1 min read

Install JetPack Without NVIDIA SDK Manager

Setting up SDK Manager is painful. Even though it’s just one single RPM or DEB package, the problem is that you...
Tanvir Kour
1 min read

Is NPU better than GPU?

When discussing hardware acceleration for AI workloads, both Neural Processing Units (NPUs) and Graphics Processing Units (GPUs) are leading technologies. However,...
Adesoji Alu
4 min read

Exploring the Revolutionary Nemotron-4-340B-Instruct: Enhanced Instruction Following and Mathematical Reasoning

Model Overview Nemotron-4-340B-Instruct is a large language model developed by NVIDIA, designed for English-based single and multi-turn chat applications. It has...
Adesoji Alu
4 min read

Running Ollama with Nvidia GPU Acceleration: A Docker Compose Guide

Discover how to harness the power of Nvidia GPUs to optimize Large Language Models like Ollama with Docker Compose in this...
Ajeet Raina
3 min read

With Grove Base Hat and OLED I2C Board for NVIDIA Jetson Nano

Grove Shield for Jetson Nano is an expansion board for Jetson Nano, designed by Seeed Studio, for the orderliness of your...
Ajeet Raina
5 min read

Docker, Prometheus & Pushgateway for NVIDIA GPU Metrics & Monitoring

In my last blog post, I talked about how to get started with NVIDIA docker & interaction with NVIDIA GPU system. I...
Ajeet Raina
3 min read

Running NVIDIA Docker in the GPU-Accelerated Data Center

Docker is the leading container platform which provides both hardware and software encapsulation by allowing multiple containers to run on the same...
Ajeet Raina
5 min read
Join our Discord Server