"You need a 24 GB GPU for serious local LLMs in 2026."
Everyone repeats this. It's not true anymore.
Just ran a 35B-parameter model on an RTX 4060 Ti 8 GB: • 41 tok/s at 16k context • 24 tok/s at 200k context
Recipe + benchmarks below 🧵 ...
Staffers for Seattle Socialist Mayor Katie Wilson abruptly end an interview with KOMO News Senior Reporter Chris Daniels when she can't answer basic questions
Wilson has been criticized for dodging the press & being unable to answer basic questions since she came into office ...
3 months, 1 week ago
37
@BrianNorgard
Thread
“Together we destroyed the city” is not a joke. Sadly, this is our reality in Los Angeles. ...
how we implemented Moondream inference on Apple Silicon (spoiler: we don't use MLX)
⬇️ (1/N)
<a target="_blank" href="https://twitter.com/mayfer/status/2050323883950313980" color="blue">x.com/mayfer/status/…</a>...
Introducing: Continuous Claude v4.7 (optimised for Opus 4.7)
strap in - we've got RLMs, 50% off Edits, 95% off Reads, fine-tuned models and even evolving codebases 👀
let's dive in to what's changed, what's new and what to do👇 ...
3 months, 1 week ago
22
@raulcontev
Thread
Hay una manera de predecir con un 91% de precisión si una pareja va a permanecer junta o terminar separándose:...