LLM stories
Large language models have moved from research curiosity to production infrastructure. This topic focuses on LLM APIs, fine-tuning, RAG pipelines, agent frameworks, and the economics of running models at scale.
HN discussions here tend to be unusually practical — benchmark comparisons, cost breakdowns, and real deployment war stories from teams shipping LLM features.
158 stories archived · Page 3 of 6
RSS feed for LLMAI Meter – Local token usage with energy and water estimates
ai-meter.app · 2 points · 2 comments
GR: Ban LLM Contributions from Debian
lists.debian.org · 6 points · 0 comments
Falsities of LLM Negationists
xlii.space · 5 points · 1 comments
GPL Infection of LLMs?
4 points · 0 comments
Debian launches competing General Resolutions on LLM usage in Debian code
debian.org · 10 points · 1 comments
AMD and Cerebras Launch AI Inference Solution
cerebras.ai · 16 points · 3 comments
Pizauth: An OAuth2 token requester daemon
github.com · 4 points · 0 comments
Why call it "inference" instead of just "model hosting"?
6 points · 9 comments
Sourceminder.org - token-efficient code search
sourceminder.org · 4 points · 2 comments
Is anyone giving out tokens for benchmarking LLMs?
3 points · 1 comments
China Wields Its Rare Earth Leverage over Europe with New Export Controls
nytimes.com · 6 points · 2 comments
A production-grade OCR pipeline on Kubernetes with vLLM and Rust
github.com · 6 points · 0 comments
How are you getting decent front end interface results out of LLMs?
2 points · 1 comments
Why I joined Blacksmith to work on storage again
blacksmith.sh · 6 points · 0 comments
Extension that shows complaints and more on NYC apartment listings
streetleaky.com · 2 points · 0 comments
My security camera shipped a GitHub admin token in its login page
hhh.hn · 25 points · 6 comments
LLM Proxy - Python with SSE stream aggregation and timeout prevention
github.com · 2 points · 0 comments
Hetzner is working on LLM Inference
sliplane.io · 51 points · 20 comments
RTK and Claude Code Token Savings: A Closer Look
blog.jetbrains.com · 5 points · 0 comments
filtersql – a dependency-free JSON-to-SQL compiler for LLMs and WebAPIs
github.com · 3 points · 1 comments
Naming a pattern in AI-generated code: Modular Mirage
2 points · 0 comments
Computer science enrollment fell for the first time in 20 years
businessinsider.com · 10 points · 3 comments
A Pragmatic Approach to LLMs
gracefulliberty.com · 6 points · 0 comments
Gigatoken-rb – A Ruby port of gigatoken that's 1.6x faster
github.com · 3 points · 0 comments
Coypa is a fun way to manage your clipboard
github.com · 3 points · 1 comments
Letterpaths: Or how LLMs can be good even when they're bad
robinlinacre.com · 4 points · 1 comments
Turo – An Aggressive Token-Saving Proxy for CLI AI Agents
github.com · 3 points · 2 comments
AI bet goes awry: Oracle fires 21,000 employees
jpost.com · 16 points · 3 comments
Connecting my dumb garage and car's homelink buttons to Home Assistant
dsauerbrun.com · 10 points · 0 comments
Ask: Why does everything Microsoft create lately feel fragile and half-baked?
9 points · 10 comments