LLM stories
Large language models have moved from research curiosity to production infrastructure. This topic focuses on LLM APIs, fine-tuning, RAG pipelines, agent frameworks, and the economics of running models at scale.
HN discussions here tend to be unusually practical — benchmark comparisons, cost breakdowns, and real deployment war stories from teams shipping LLM features.
158 stories archived · Page 2 of 6
RSS feed for LLMWattage: A token-spend profiler and cost-regression gate for AI agents
github.com · 4 points · 0 comments
What's the best hands-on path to learn ML inference infrastructure?
5 points · 2 comments
An interactive way to exploit an LLM without going to jail
github.com · 2 points · 0 comments
Anyone else experiencing extreme Codex limits token burn rate?
2 points · 1 comments
The relay market powering token resellers and fraud
vectoral.com · 132 points · 84 comments
Hallmark – Anti-AI-Slop Design Skill for Claude Code, Cursor, and Codex
github.com · 6 points · 8 comments
CrispVoice – Studio voice enhancement that never uploads your voice
github.com · 5 points · 1 comments
Claude Code Cut Their System Prompt by 80%. Does That Work for Small Models Too?
antigma.ai · 5 points · 4 comments
VoiceScroll – a teleprompter that scrolls as you speak
voice-scroll.com · 3 points · 0 comments
I built a hypervisor and client for inference on consumer compute
scalattice.com · 2 points · 0 comments
Agentic test processes, LLM benchmarks, and other notes on agentic coding
danluu.com · 7 points · 0 comments
"Loneliness influencers" are all the rage, but everyone is wrong about them
preta6.substack.com · 5 points · 0 comments
GM Backs Sodium Ion Batteries for U.S. Grid Storage
spectrum.ieee.org · 159 points · 61 comments
Becoming a Research Engineer at a Big LLM Lab
maxmynter.com · 26 points · 10 comments
Awsmux – Multi-account AWS CLI, up to 5.4x faster, 7.4x fewer tokens
github.com · 7 points · 4 comments
LLM Usage in Debian: Three Proposals
debian.org · 114 points · 104 comments
Running a 28.9M parameter LLM on an $8 microcontroller
github.com · 120 points · 27 comments
What is the status on continual learning for LLMs?
5 points · 13 comments
HotPin – lossless 120B MoE inference on 24GB RAM (CPU, 50 loc)
7 points · 0 comments
A system prompt to get AI to stop pretending to be human
swiftrocks.com · 31 points · 13 comments
Politician reads AI prompt during assembly
youtube.com · 67 points · 43 comments
Brazilian farmers tokenized dairy cows to get loans, bypassing bank limits
coindesk.com · 59 points · 49 comments
The AI Boom Made Average People More Interesting
alec.is · 6 points · 0 comments
Will it be possible to do LLM Pool Training?
2 points · 2 comments
What happens behind the scenes when we change effort for same LLM models?
11 points · 8 comments
Has anyone figured out how to guard LLMs on kernel level?
2 points · 1 comments
TS Compiler Knowledge Graph reducing AI tokens about 90%
github.com · 3 points · 0 comments
The Productivity Mirage
frantic.im · 11 points · 0 comments
Opposing ICE Might Save the Country. It Could Also Ruin Your Life
wired.com · 7 points · 2 comments
"We removed over 80% of Claude Code's system prompt for Opus 5 and Fable 5"
twitter.com · 3 points · 0 comments