AIToday
Large Language ModelsOpen-Source AIr/LocalLLaMAPublished: Apr 20, 2026, 22:00 JST1 min read

AMD enthusiast builds local LLM workstation achieving 120 tokens/second with Ryzen 9700X and Radeon R9700, seeking optimal model recommendations.

3 Key Points

  1. Builder configured high-end local inference rig with AMD Radeon AI PRO R9700 (32GB VRAM) and Ryzen 7 9700X CPU paired with 64GB DDR5 RAM

  2. System achieves ~120 tokens/second performance on simple prompts using Qwen3.6-35B-A3B model via LM Studio with Vulkan backend

  3. Poster seeks community advice on largest compatible model architectures and whether Q4_K_M quantizations are optimal for their hardware setup

Ask the AI about this article →

Get the latest Large Language Models news every morning

For example, today's edition would include:

  • Anthropic releases Claude Fable 5.1 and Mythos 5.1ITmedia AI+ · 2h ago
  • LLM serving: why continuous batching winsDaily Dose of Data Science · 2h ago
  • Anthropic's Claude Fable 5.1 Now on Snowflake Cortex AISnowflake AI Blog · 2h ago

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · takes 30 seconds · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. Q&As are published on this page for other readers too.

Related Articles

Next articleResearch suggests 50% AI-assisted writing hits the optimal balance between human authenticity and AI efficiency