All topics

LLM stories

Large language models have moved from research curiosity to production infrastructure. This topic focuses on LLM APIs, fine-tuning, RAG pipelines, agent frameworks, and the economics of running models at scale.

HN discussions here tend to be unusually practical — benchmark comparisons, cost breakdowns, and real deployment war stories from teams shipping LLM features.

158 stories archived · Page 2 of 6

RSS feed for LLM

Wattage: A token-spend profiler and cost-regression gate for AI agents

github.com · 4 points · 0 comments

What's the best hands-on path to learn ML inference infrastructure?

5 points · 2 comments

An interactive way to exploit an LLM without going to jail

github.com · 2 points · 0 comments

Anyone else experiencing extreme Codex limits token burn rate?

2 points · 1 comments

The relay market powering token resellers and fraud

vectoral.com · 132 points · 84 comments

Hallmark – Anti-AI-Slop Design Skill for Claude Code, Cursor, and Codex

github.com · 6 points · 8 comments

CrispVoice – Studio voice enhancement that never uploads your voice

github.com · 5 points · 1 comments

Claude Code Cut Their System Prompt by 80%. Does That Work for Small Models Too?

antigma.ai · 5 points · 4 comments

VoiceScroll – a teleprompter that scrolls as you speak

voice-scroll.com · 3 points · 0 comments

I built a hypervisor and client for inference on consumer compute

scalattice.com · 2 points · 0 comments

Agentic test processes, LLM benchmarks, and other notes on agentic coding

danluu.com · 7 points · 0 comments

"Loneliness influencers" are all the rage, but everyone is wrong about them

preta6.substack.com · 5 points · 0 comments

GM Backs Sodium Ion Batteries for U.S. Grid Storage

spectrum.ieee.org · 159 points · 61 comments

Becoming a Research Engineer at a Big LLM Lab

maxmynter.com · 26 points · 10 comments

Awsmux – Multi-account AWS CLI, up to 5.4x faster, 7.4x fewer tokens

github.com · 7 points · 4 comments

LLM Usage in Debian: Three Proposals

debian.org · 114 points · 104 comments

Running a 28.9M parameter LLM on an $8 microcontroller

github.com · 120 points · 27 comments

What is the status on continual learning for LLMs?

5 points · 13 comments

HotPin – lossless 120B MoE inference on 24GB RAM (CPU, 50 loc)

7 points · 0 comments

A system prompt to get AI to stop pretending to be human

swiftrocks.com · 31 points · 13 comments

Politician reads AI prompt during assembly

youtube.com · 67 points · 43 comments

Brazilian farmers tokenized dairy cows to get loans, bypassing bank limits

coindesk.com · 59 points · 49 comments

The AI Boom Made Average People More Interesting

alec.is · 6 points · 0 comments

Will it be possible to do LLM Pool Training?

2 points · 2 comments

What happens behind the scenes when we change effort for same LLM models?

11 points · 8 comments

Has anyone figured out how to guard LLMs on kernel level?

2 points · 1 comments

TS Compiler Knowledge Graph reducing AI tokens about 90%

github.com · 3 points · 0 comments

The Productivity Mirage

frantic.im · 11 points · 0 comments

Opposing ICE Might Save the Country. It Could Also Ruin Your Life

wired.com · 7 points · 2 comments

"We removed over 80% of Claude Code's system prompt for Opus 5 and Fable 5"

twitter.com · 3 points · 0 comments