🧠 Inference is the next phase of AI, and inference is bottlenecked...

@RosannaInvests
Rosanna Prestia, MBA@RosannaInvests
35 views Jun 22, 2026 ~1 min read
Advertisement
1
🧠 Inference is the next phase of AI, and inference is bottlenecked by memory, not compute.

Every agent, every model serving real users, needs fast memory sitting next to the GPU. The market is starting to price the GPUs. It has not yet priced the memory layer underneath them.

$PENG (Penguin Solutions) sits exactly there. 🐧 Thread. 🧡
Media image
2
The thesis in one line.

As AI shifts from training to inference, the constraint moves from raw compute to memory bandwidth and capacity. You cannot serve a model fast if the data cannot reach the processor fast.

$PENG builds the memory architecture for that exact problem β†’ CXL memory, KV cache servers, and a photonic memory appliance in development.

This is the inference-memory chokepoint. And it is profitable today. πŸ’Έ
Actions
What You Can Do
  • Export as PDF or Markdown
  • Batch Export to Notion
  • Bookmark & Highlight
  • LinkedIn & Instagram Carousel Maker
Create Free Account

Includes 7-day Premium trial

Advertisement