All topics

LLM stories

Large language models have moved from research curiosity to production infrastructure. This topic focuses on LLM APIs, fine-tuning, RAG pipelines, agent frameworks, and the economics of running models at scale.

HN discussions here tend to be unusually practical — benchmark comparisons, cost breakdowns, and real deployment war stories from teams shipping LLM features.

158 stories archived · Page 1 of 6

RSS feed for LLM

Vyne – Zapier for DeFi users, run from any LLM or the workflow builder

app.vyne.finance · 2 points · 0 comments

I sent Claude Opus 5 '–-' and it wrote me 5k tokens about a cartographer

austinsnerdythings.com · 15 points · 5 comments

OpenReviewer: A Specialized LLM for Generating Critical Scientific Paper Reviews

aclanthology.org · 4 points · 0 comments

What's Wrong with American Studies?

radicallypragmatic.org · 5 points · 1 comments

What if useful AI is a fantasy?

lzon.ca · 9 points · 3 comments

Writekin – fine-tune a local LLM on your own writing, on your Mac

github.com · 2 points · 0 comments

"Uncensored" open LLMs are measurably more optimistic than their base models

arxiv.org · 23 points · 7 comments

SOTA on the hardest AI memory benchmark (BEAM, 10M tokens), with a smaller model

2 points · 3 comments

Now Is the Time to Give LLMs Access to the ACM Digital Library

cacm.acm.org · 85 points · 62 comments

Ctrlb-decompose: Strip the noise from logs before sending to LLMs

github.com · 48 points · 6 comments

Claude Code system prompt says not to ask for permission, assumes user is absent

elliotmilco.substack.com · 4 points · 0 comments

Mondragon Corporation – a federation of co-operatives

en.wikipedia.org · 161 points · 31 comments

DMARC has been public since 2012 but most company domains still don't enforce it

ciphercue.com · 160 points · 100 comments

Does Speaking to Agents Like Cavemen Save 65% of Tokens? We Test

blog.jetbrains.com · 5 points · 1 comments

Orchard – Let AI agents set up your app's back end with one prompt

orchard-dashboard.pages.dev · 2 points · 0 comments

Day 0 Kimi-K3 Inference Deployment with Atom on AMD Instinct MI355X GPUs

amd.com · 9 points · 0 comments

PromptTrace – Free hands-on labs to practice hacking LLMs

prompttrace.airedlab.com · 3 points · 0 comments

Don't ask an LLM for a confidence score

justinflick.com · 20 points · 1 comments

Measured LLM inference speeds on Apple Silicon, with raw data (CC BY 4.0)

macyou.co · 10 points · 3 comments

Kimi K3 Now Available via Telnyx Inference API

telnyx.com · 69 points · 24 comments

Professor's invisible prompt trap catches 32/35 students cheating with AI

techspot.com · 81 points · 74 comments

Evading Residential Proxy Networks

fbi.gov · 4 points · 2 comments

Kimi K3 on vLLM: Up to 370 Tokens/sec

vllm.ai · 5 points · 0 comments

Ctxdiff – Git diff for your LLM agent's context window

github.com · 3 points · 3 comments

Where do those LLM-generated outreach emails come from?

3 points · 1 comments

General Resolution: LLM Usage in Debian

debian.org · 3 points · 0 comments

ASD-STE100 Simplified Technical English for LLMs

github.com · 5 points · 2 comments

Are we having a substantial increase of "Show HN" since LLM aided coding

7 points · 4 comments

Sand battery: Finland's answer to a renewable energy headache

cnbc.com · 27 points · 8 comments

Aqua UI for the Web

github.com · 2 points · 0 comments