view post Post 33 Real Long-Term Memory for AI: A 50-Million-Token Window That Is Faster and Cheaper Than Recompute https://www.reddit.com/r/huggingface/s/T2tXxKlqSb See translation 👀 1 1 + Reply
view post Post 4119 You can now train your own Decision model like Jev locally!We increased Qwen3.5 0.8B’s aggregate accuracy from 20.7% to 74.3% across 3 decision benchmarks - on just 4GB VRAM.Turn any LLM like Qwen3.8, Gemma 4 into decision models with our open-source Unsloth repo.GitHub: https://github.com/unslothai/unslothGuide: https://unsloth.ai/docs/basics/train-your-own-decision-model-with-unsloth See translation 4 replies · ❤️ 17 17 🔥 8 8 👍 4 4 ➕ 2 2 🧠 2 2 + Reply
view post Post 33 Real Long-Term Memory for AI: A 50-Million-Token Window That Is Faster and Cheaper Than Recompute https://www.reddit.com/r/huggingface/s/T2tXxKlqSb See translation 👀 1 1 + Reply
Real Long-Term Memory for AI: A 50-Million-Token Window That Is Faster and Cheaper Than Recompute Paper • 2610.10845 • Published 4 days ago • 8 • 2
Real Long-Term Memory for AI: A 50-Million-Token Window That Is Faster and Cheaper Than Recompute Paper • 2610.10845 • Published 4 days ago • 8
Real Long-Term Memory for AI: A 50-Million-Token Window That Is Faster and Cheaper Than Recompute Paper • 2610.10845 • Published 4 days ago • 8
Working Around the Compute Ceiling: Byte-Exact Memory in Galahad Makes LLM Reading a One-Time Cost LLM Reading a One-Time Cost Paper • 2609.39358 • Published 11 days ago • 4
Working Around the Compute Ceiling: Byte-Exact Memory in Galahad Makes LLM Reading a One-Time Cost LLM Reading a One-Time Cost Paper • 2609.39358 • Published 11 days ago • 4
A Frozen 12B Beats Frontier Models on Verified Work: 100% Accuracy, 0 Tokens, Bit-Exact, Forever Paper • 2607.23806 • Published Jul 26 • 8
Smarter and Cheaper at Once: Byte-Exact KV-Cache Grafting Turns a Frozen Small Model into a Verified-Knowledge Flywheel Paper • 2607.14431 • Published Jul 15 • 7
A Frozen 12B Beats Frontier Models on Verified Work: 100% Accuracy, 0 Tokens, Bit-Exact, Forever Paper • 2607.23806 • Published Jul 26 • 8 • 4
A Frozen 12B Beats Frontier Models on Verified Work: 100% Accuracy, 0 Tokens, Bit-Exact, Forever Paper • 2607.23806 • Published Jul 26 • 8
A Frozen 12B Beats Frontier Models on Verified Work: 100% Accuracy, 0 Tokens, Bit-Exact, Forever Paper • 2607.23806 • Published Jul 26 • 8
Running Agents A 12B Model Beats Fable 5 at 0 Tokens ⚡ 180/180 at 0 tokens, bit-exact, on verified work.
Running Agents A 12B Model Beats Fable 5 at 0 Tokens ⚡ 180/180 at 0 tokens, bit-exact, on verified work.
Smarter and Cheaper at Once: Byte-Exact KV-Cache Grafting Turns a Frozen Small Model into a Verified-Knowledge Flywheel Paper • 2607.14431 • Published Jul 15 • 7
Smarter and Cheaper at Once: Byte-Exact KV-Cache Grafting Turns a Frozen Small Model into a Verified-Knowledge Flywheel Paper • 2607.14431 • Published Jul 15 • 7 • 7
Smarter and Cheaper at Once: Byte-Exact KV-Cache Grafting Turns a Frozen Small Model into a Verified-Knowledge Flywheel Paper • 2607.14431 • Published Jul 15 • 7 • 7