Thread Navigator
  • Home
  • Explore
  • Videos
    • Carousel Maker
    • Screenshot Maker
  • AI Agents
  • Log In
  • Sign Up
  • Home
  • Explore
  • Videos
  • Create
  • Carousel Maker
  • Screenshot Maker
  • AI Agents
@navaneethvb

Navaneeth Krishnan (@navaneethvb)

View on X 1 Unrolled Thread

Topics they write about

🤖 AI & Machine Learning 1
Thread Archive
13
🤖 AI & Machine Learning

Apparently, one way to make LLM inference almost 2x faster is to make the CPU stop talking to the GPU so much. And this is where CUDA graphs come in, but before getting into the details let us actually understand the problem first. An LLM forward pass is not one giant GPU operation. Underneath it,...

Aug 23, 2026
1
Thread Navigator
AboutPrivacyTermsDisclaimer
Made with ❤️ for the reading community ·v2.7.0