Lenovo ThinkStation PGX: Mini Computer Runs 200B AI Models

Technology Artificial Intelligence

Aug 15, 2026 · 4 min read

Lenovo ThinkStation PGX: Mini Computer Runs 200B AI Models

The Lenovo ThinkStation PGX is a powerful mini computer that can run AI models with up to 200 billion parameters, making it ideal for both professional and personal use in tight spaces. Its compact size belies impressive specs, including a 1 petaflop performance and high-bandwidth connectivity options.

Source

Watch the Reel

AI Models on Mini Computer with Lenovo ThinkStation PGX

Mini computers have become increasingly popular for their compact size and powerful capabilities. The Lenovo ThinkStation PGX is a standout in this category, offering impressive AI model processing capabilities in a small form factor.

Why This Matters

The Lenovo ThinkStation PGX is designed to pack a punch despite its small size. Its ability to run large AI models locally makes it a strong contender for both professional and personal use. Understanding its capabilities and limitations can help users make informed decisions about integrating this device into their tech ecosystems.

Key Features of the Lenovo ThinkStation PGX

Hardware Specifications

The Lenovo ThinkStation PGX is equipped with the Nvidia GB10 Grace Blackwell Superchip and 128GB of unified memory. This powerful combination allows the device to handle AI models with up to 200 billion parameters. The device's small form factor, smaller than a Mac Mini, makes it an attractive option for those with limited desk space.

Performance

The PGX boasts a 1 petaflop FP4 performance, which is a significant achievement for a device of its size. This high performance is further enhanced by its USB-C power delivery at 240W, ensuring that the device operates efficiently without requiring a bulky power supply.

The Connectivity Trick

High-Bandwidth Device-to-Device Connections

One of the standout features of the Lenovo ThinkStation PGX is its dual QSFP ports, facilitated by a ConnectX-7 NIC. These ports enable 200 GbE direct device-to-device connections, which are crucial for high-bandwidth workloads. This feature is particularly useful for distributed computing tasks, as it allows two PGX units to link and work together seamlessly using frameworks like NCCL, Megatron-LM, or vLLM multi-node.

Memory Limitations

While the connectivity options are robust, it's important to note that the PGX does not create a unified 256GB shared memory pool. Memory coherency is limited to the system boundary, meaning that while you get high-bandwidth distributed computing, you won't have hardware-level memory pooling. This is a critical consideration for users who need extensive memory sharing across multiple devices.

Benchmark Performance

196B Mixture-of-Experts Model

The Lenovo ThinkStation PGX has been benchmarked to run a 196B parameter Mixture-of-Experts (MoE) model efficiently. Using the Step-3.5-Flash model in Q4_K_S GGUF format via llama.cpp, the PGX achieved approximately 20 tokens per second at 50ms latency. This performance showcases the unified memory advantage, allowing community tools like llama.cpp to leverage the full 128GB memory pool directly. This capability is unmatched by discrete GPUs at a similar price point.

Unified Memory Advantage

The unified memory architecture of the PGX is a significant advantage. It allows for efficient memory utilization, ensuring that large AI models can run smoothly without the need for extensive memory management. This is particularly beneficial for tasks that require high memory bandwidth and low latency.

Practical Tips

Optimizing Performance

To get the most out of the Lenovo ThinkStation PGX, consider the following tips:

  1. Leverage Unified Memory: Make use of the full 128GB memory pool for large AI models. This will ensure that your tasks run smoothly and efficiently.
  2. Utilize Connectivity Features: If you need to distribute workloads across multiple devices, take advantage of the dual QSFP ports for high-bandwidth connections.
  3. Monitor Memory Coherency: Be aware of the memory coherency limitations and plan your workloads accordingly to avoid bottlenecks.

Troubleshooting Common Issues

  1. Power Management: Ensure that the device is connected to a reliable power source. The 240W USB-C power delivery is efficient but may require a high-quality cable.
  2. Connectivity Problems: If you encounter issues with device-to-device connections, check the QSFP ports and the ConnectX-7 NIC for any potential faults.
  3. Memory Utilization: If your tasks are memory-intensive, monitor the memory usage and consider optimizing your code for better performance.

Important Takeaways

The Lenovo ThinkStation PGX is a powerful mini computer designed for AI model processing. Its key features, including the Nvidia GB10 Grace Blackwell Superchip, 128GB unified memory, and dual QSFP ports, make it a strong contender in the mini computer market. However, it's important to be aware of its memory limitations and understand how to optimize its performance for your specific needs.

Conclusion

The Lenovo ThinkStation PGX is a game-changer in the world of mini computers, offering robust AI processing capabilities in a compact form factor. Whether you're a professional looking to run large AI models locally or a tech enthusiast interested in high-performance computing, the PGX has a lot to offer. By understanding its features, limitations, and best practices, you can make the most of this powerful device and integrate it seamlessly into your tech ecosystem.

Answers

FAQ

The Lenovo ThinkStation PGX is designed to handle AI models with up to 200 billion parameters, supported by a 1 petaflop performance and high-bandwidth connectivity options. This makes it capable of running complex AI tasks efficiently, even in a small form factor.

Mentioned

Products

mini computer
Discussion

Comments

Be the first to comment.

Similar reads based on topic and creator.

Recent articles

Fresh deep dives from the latest Reels we unpacked.

View all