Tencent

UI Mate 9B

Tencent model discovered on huggingface

Model Summary

Family

UI

Version

9

Parameters

9.4B (est.)

Parameter counts for closed models are estimates; vendors rarely publish exact sizes.

VRAM Requirements by Quantization

Memory needed to serve UI Mate 9B for inference, including a 20% overhead for activations and KV cache.

PrecisionVRAM neededSmallest single GPU that fits
INT4 (4-bit)5.25 GBNVIDIA P100 SXM2 (16 GB)
INT8 (8-bit)10.51 GBNVIDIA P100 SXM2 (16 GB)
FP16 (16-bit)21.01 GBNVIDIA L4 (24 GB)
FP32 (32-bit)42.02 GBNVIDIA L40S (48 GB)

Recommended GPU Configurations

Cheapest on-demand configurations to serve UI Mate 9B at 8-bit (11 GB VRAM).

1x NVIDIA T4

16 GB total VRAM · Turing

~$0.50/h

1x NVIDIA P100 SXM2

16 GB total VRAM · Pascal

~$0.60/h

1x NVIDIA L4

24 GB total VRAM · Ada Lovelace

~$1.00/h

Quick GPU Planning

Use the calculator pre-filled with this exact version to estimate memory, speed, and compute requirements in a few clicks.

Access Pre-filled Calculator

Frequently Asked Questions

How much VRAM do you need to run UI Mate 9B?

With an estimated 9.4B parameters, UI Mate 9B needs roughly 11 GB of VRAM in 8-bit (INT8), 5 GB in 4-bit, and 21 GB in FP16, including a 20% overhead for activations and KV cache.

Which GPUs can run UI Mate 9B?

At 8-bit quantization, the most cost-effective option is 1x NVIDIA T4 (16 GB combined VRAM, around $0.50/hour on-demand). Higher-end cards like the NVIDIA B200 or AMD MI355X reduce the GPU count needed.

Can UI Mate 9B run on a single GPU?

Yes. In 8-bit, a single NVIDIA P100 SXM2 (16 GB) fits the model.

How much does it cost to serve UI Mate 9B in the cloud?

Renting 1x T4 costs on the order of $0.50/hour, i.e. about $365/month running 24/7. Actual prices vary by provider and commitment; spot and reserved capacity can be significantly cheaper.

Other UI Versions