133 shaares
1 result
tagged
qwen
Using an HP Workstation Z2 Mini G3 equipped with an Intel Core i7–7700, 16GB of RAM, and an NVIDIA Quadro M620 with just 2GB of VRAM, I explored how far modern quantized language models can be pushed using llama.cpp.