LLM stories
Large language models have moved from research curiosity to production infrastructure. This topic focuses on LLM APIs, fine-tuning, RAG pipelines, agent frameworks, and the economics of running models at scale.
HN discussions here tend to be unusually practical — benchmark comparisons, cost breakdowns, and real deployment war stories from teams shipping LLM features.
442 stories archived · Page 1 of 15
RSS feed for LLMMicron, SK Commit Billions to RAM Capacity, but Almost Nothing Lands Before 2028
storagereview.com · 21 points · 3 comments
Declarative-forms – await an object the way prompt() awaits a string
wolfoo2931.github.io · 4 points · 0 comments
If you're burning tokens at scale, what are you using them for?
4 points · 3 comments
LLM City – 3D render of all Kimi K3's weights as 2.5mm tiles
magik.net · 9 points · 1 comments
Universal Health Coverage Could Save $1T and 114,000 Lives a Year, Yale Study
ysph.yale.edu · 76 points · 46 comments
LLMs each trading $100K vs. a frozen rulebook – the rulebook leads
aitradingcompetition.com · 4 points · 0 comments
When do you think LLM capacity will reach its ceiling?
2 points · 0 comments
What LLM subscription/provider to use with pi harness?
3 points · 2 comments
Sib - Unixy LLM Client using Git to store converastions, instead SQLite
github.com · 4 points · 0 comments
The AI Credit Resale Economy
vectoral.com · 203 points · 76 comments
Claude: System Prompts
platform.claude.com · 472 points · 207 comments
Harness session memory without transcript hoarding
llm-wiki.net · 4 points · 0 comments
What happens when an LLM never sees material beyond fifth grade?
littlelearner-ll.github.io · 243 points · 208 comments
Widen, a native Postgres GUI using Apple's on-device LLM
github.com · 9 points · 0 comments
Fairly Ranking the Most Brilliant Birds
moultano.wordpress.com · 3 points · 0 comments
It's How You Ask: Gender-Associated Linguistic Bias in LLMs
arxiv.org · 16 points · 7 comments
California Energy Storage System Survey
energy.ca.gov · 3 points · 1 comments
Between Tokens – an interactive piece where you are the language model
chrisjz.github.io · 5 points · 2 comments
DeepSeek V4 Flash at 278 tok/s, full precision, no quantization
runinfra.ai · 5 points · 3 comments
Debian has begun voting on the future of AI/LLM contributions
lists.debian.org · 9 points · 1 comments
Suspecting court of using AI, man injected prompts in filings to try to win case
arstechnica.com · 24 points · 17 comments
ThoughtDAG – An editable context graph for LLM conversations
chenxiachan.github.io · 37 points · 12 comments
destruction-certificate.txt
stallman.org · 39 points · 1 comments
Dictata – Local Whisper dictation with LLM cleanup
github.com · 2 points · 0 comments
Baking a Model: A Metaphor for LLM Training
newsletter.kentbeck.com · 5 points · 0 comments
Kvcachescope – Why Nvidia-smi is blind to vLLM KV cache leaks
github.com · 3 points · 0 comments
Being Against LLMs Is Against the Spirit of Floss
joarvarndt.se · 13 points · 10 comments
Just how big is the hidden leverage of AI hyperscalers?
ft.com · 7 points · 1 comments
A Contract-Grade Verifier for LLM-Generated GPU Kernels
arxiv.org · 14 points · 0 comments
Graft – Claude Code hooks that cut grep tokens by 42%
github.com · 36 points · 36 comments