This guide documents the process to optimize and run large language models on AMD APUs (specifically 7840U with Radeon 780M, Ayaneo Flip w/ 32GB RAM) using llama.cpp with Vulkan acceleration.
# Disable read-only filesystem (SteamOS specific)
sudo steamos-readonly disable