Multi-Instance GPU

Multi-Instance GPU (MIG) is an NVIDIA technology that partitions a single GPU into several isolated instances.

1 article
Last mentioned

Multi-Instance GPU (MIG) is a technology NVIDIA introduced in 2020 with the A100, its first Ampere-architecture data center GPU. It divides a GPU's compute units and memory into as many as seven instances, each of which behaves like a separate GPU.

Data centers and cloud platforms use it to share one GPU among multiple users or workloads. Because each instance receives only part of the GPU's resources, it can run the same task more slowly than the full GPU.

This entry is based on AIPOST articles and widely known facts. If something is wrong, please send us a correction request.

Articles covering this entry

…GPUs can reduce assignment delays, but testing found inference on a Blackwell MIG slice—a portion of a GPU allocated as a separate instance—was 1.8 times slower…


© 2026 AIPOST. All rights reserved.

AIPOST is an AI publication covering practical AI, AI security, performance, startups, health, ethics and industry news. No account is needed, and our privacy policy explains how we handle personal information.