All topics

LLM stories

Large language models have moved from research curiosity to production infrastructure. This topic focuses on LLM APIs, fine-tuning, RAG pipelines, agent frameworks, and the economics of running models at scale.

HN discussions here tend to be unusually practical — benchmark comparisons, cost breakdowns, and real deployment war stories from teams shipping LLM features.

158 stories archived · Page 3 of 6

RSS feed for LLM

AI Meter – Local token usage with energy and water estimates

ai-meter.app · 2 points · 2 comments

GR: Ban LLM Contributions from Debian

lists.debian.org · 6 points · 0 comments

Falsities of LLM Negationists

xlii.space · 5 points · 1 comments

GPL Infection of LLMs?

4 points · 0 comments

Debian launches competing General Resolutions on LLM usage in Debian code

debian.org · 10 points · 1 comments

AMD and Cerebras Launch AI Inference Solution

cerebras.ai · 16 points · 3 comments

Pizauth: An OAuth2 token requester daemon

github.com · 4 points · 0 comments

Why call it "inference" instead of just "model hosting"?

6 points · 9 comments

Sourceminder.org - token-efficient code search

sourceminder.org · 4 points · 2 comments

Is anyone giving out tokens for benchmarking LLMs?

3 points · 1 comments

China Wields Its Rare Earth Leverage over Europe with New Export Controls

nytimes.com · 6 points · 2 comments

A production-grade OCR pipeline on Kubernetes with vLLM and Rust

github.com · 6 points · 0 comments

How are you getting decent front end interface results out of LLMs?

2 points · 1 comments

Why I joined Blacksmith to work on storage again

blacksmith.sh · 6 points · 0 comments

Extension that shows complaints and more on NYC apartment listings

streetleaky.com · 2 points · 0 comments

My security camera shipped a GitHub admin token in its login page

hhh.hn · 25 points · 6 comments

LLM Proxy - Python with SSE stream aggregation and timeout prevention

github.com · 2 points · 0 comments

Hetzner is working on LLM Inference

sliplane.io · 51 points · 20 comments

RTK and Claude Code Token Savings: A Closer Look

blog.jetbrains.com · 5 points · 0 comments

filtersql – a dependency-free JSON-to-SQL compiler for LLMs and WebAPIs

github.com · 3 points · 1 comments

Naming a pattern in AI-generated code: Modular Mirage

2 points · 0 comments

Computer science enrollment fell for the first time in 20 years

businessinsider.com · 10 points · 3 comments

A Pragmatic Approach to LLMs

gracefulliberty.com · 6 points · 0 comments

Gigatoken-rb – A Ruby port of gigatoken that's 1.6x faster

github.com · 3 points · 0 comments

Coypa is a fun way to manage your clipboard

github.com · 3 points · 1 comments

Letterpaths: Or how LLMs can be good even when they're bad

robinlinacre.com · 4 points · 1 comments

Turo – An Aggressive Token-Saving Proxy for CLI AI Agents

github.com · 3 points · 2 comments

AI bet goes awry: Oracle fires 21,000 employees

jpost.com · 16 points · 3 comments

Connecting my dumb garage and car's homelink buttons to Home Assistant

dsauerbrun.com · 10 points · 0 comments

Ask: Why does everything Microsoft create lately feel fragile and half-baked?

9 points · 10 comments