An Old Laptop’s Only M.2 Slot Was Used To Hook Up An RX 7900 XT, Achieving 60 Tokens/s In Qwen3.6 27B & 100K Context Window; External Drive Was Used For The OS

A user repurposed an old laptop's M.2 slot to connect a powerful desktop GPU, enabling local LLM inference. This setup achieved impressive speeds with a 27 billion parameter model and a 100,000 token context window, using an external drive for the operating system.
Key takeaways
- M.2 slot can host external GPUs for AI tasks
- Achieved 60 tokens/s with a 27B parameter model
- Successfully utilized a 100K context window locally
- External drive supported the OS for this setup
Why it matters
This demonstrates a creative workaround for running demanding AI models on less powerful hardware. It suggests that with ingenuity, users can bypass typical hardware limitations for local AI tasks, potentially lowering the barrier to entry for advanced LLM experimentation.
Try this on SynaBot
Related AI assistants, prompts, and tools from the SynaBot catalog.



