Thread Navigator
  • Home
  • Explore
  • Videos
    • Carousel Maker
    • Screenshot Maker
  • AI Agents
  • Log In
  • Sign Up
  • Home
  • Explore
  • Videos
  • Create
  • Carousel Maker
  • Screenshot Maker
  • AI Agents
@leftcurvedev_

left curve dev (@leftcurvedev_)

View on X 1 Unrolled Thread

Topics they write about

🤖 AI & Machine Learning 1
Thread Archive
179
🤖 AI & Machine Learning

Anyone with 8GB or 12GB VRAM setups needs to understand that "-ncmoe" is the key flag to boost performance on llama.cpp Here are my results for Qwen3.6 35B A3B, with 64k q8_0 context on a 8GB RTX 3070Ti: ⚪️ no flag → 8.7 tok/s RAM: 13.6GB & VRAM: 7.8GB 🔴 -ncmoe 35 → 27.5 tok/s RAM: 12.1GB & VRAM:...

May 08, 2026
1
Thread Navigator
AboutPrivacyTermsDisclaimer
Made with ❤️ for the reading community ·v2.7.0