All topics

LLM stories

Large language models have moved from research curiosity to production infrastructure. This topic focuses on LLM APIs, fine-tuning, RAG pipelines, agent frameworks, and the economics of running models at scale.

HN discussions here tend to be unusually practical — benchmark comparisons, cost breakdowns, and real deployment war stories from teams shipping LLM features.

442 stories archived · Page 1 of 15

RSS feed for LLM

Micron, SK Commit Billions to RAM Capacity, but Almost Nothing Lands Before 2028

storagereview.com · 21 points · 3 comments

Declarative-forms – await an object the way prompt() awaits a string

wolfoo2931.github.io · 4 points · 0 comments

If you're burning tokens at scale, what are you using them for?

4 points · 3 comments

LLM City – 3D render of all Kimi K3's weights as 2.5mm tiles

magik.net · 9 points · 1 comments

Universal Health Coverage Could Save $1T and 114,000 Lives a Year, Yale Study

ysph.yale.edu · 76 points · 46 comments

LLMs each trading $100K vs. a frozen rulebook – the rulebook leads

aitradingcompetition.com · 4 points · 0 comments

When do you think LLM capacity will reach its ceiling?

2 points · 0 comments

What LLM subscription/provider to use with pi harness?

3 points · 2 comments

Sib - Unixy LLM Client using Git to store converastions, instead SQLite

github.com · 4 points · 0 comments

The AI Credit Resale Economy

vectoral.com · 203 points · 76 comments

Claude: System Prompts

platform.claude.com · 472 points · 207 comments

Harness session memory without transcript hoarding

llm-wiki.net · 4 points · 0 comments

What happens when an LLM never sees material beyond fifth grade?

littlelearner-ll.github.io · 243 points · 208 comments

Widen, a native Postgres GUI using Apple's on-device LLM

github.com · 9 points · 0 comments

Fairly Ranking the Most Brilliant Birds

moultano.wordpress.com · 3 points · 0 comments

It's How You Ask: Gender-Associated Linguistic Bias in LLMs

arxiv.org · 16 points · 7 comments

California Energy Storage System Survey

energy.ca.gov · 3 points · 1 comments

Between Tokens – an interactive piece where you are the language model

chrisjz.github.io · 5 points · 2 comments

DeepSeek V4 Flash at 278 tok/s, full precision, no quantization

runinfra.ai · 5 points · 3 comments

Debian has begun voting on the future of AI/LLM contributions

lists.debian.org · 9 points · 1 comments

Suspecting court of using AI, man injected prompts in filings to try to win case

arstechnica.com · 24 points · 17 comments

ThoughtDAG – An editable context graph for LLM conversations

chenxiachan.github.io · 37 points · 12 comments

destruction-certificate.txt

stallman.org · 39 points · 1 comments

Dictata – Local Whisper dictation with LLM cleanup

github.com · 2 points · 0 comments

Baking a Model: A Metaphor for LLM Training

newsletter.kentbeck.com · 5 points · 0 comments

Kvcachescope – Why Nvidia-smi is blind to vLLM KV cache leaks

github.com · 3 points · 0 comments

Being Against LLMs Is Against the Spirit of Floss

joarvarndt.se · 13 points · 10 comments

Just how big is the hidden leverage of AI hyperscalers?

ft.com · 7 points · 1 comments

A Contract-Grade Verifier for LLM-Generated GPU Kernels

arxiv.org · 14 points · 0 comments

Graft – Claude Code hooks that cut grep tokens by 42%

github.com · 36 points · 36 comments