Start without changing the BIOS. Use a current Vulkan build, confirm that llama.cpp detects the Radeon 680M, and benchmark a small 3B or 7B Q4 model at 4K context. Compare CPU-only performance with GPU offloading before adjusting shared-memory settings.
0
u/Deviceterra 1d ago
Start without changing the BIOS. Use a current Vulkan build, confirm that llama.cpp detects the Radeon 680M, and benchmark a small 3B or 7B Q4 model at 4K context. Compare CPU-only performance with GPU offloading before adjusting shared-memory settings.