Using an HP Workstation Z2 Mini G3 equipped with an Intel Core i7–7700, 16GB of RAM, and an NVIDIA Quadro M620 with just 2GB of VRAM, I explored how far modern quantized language models can be pushed using llama.cpp.
llm
llamacpp
ai
qwen
deepseek
Nvidia
FoldFold allExpandExpand allAre you sure you want to delete this link?Are you sure you want to delete this tag?
The personal, minimalist, super fast, database-free, bookmarking service by the Shaarli community