LLM stories
Large language models have moved from research curiosity to production infrastructure. This topic focuses on LLM APIs, fine-tuning, RAG pipelines, agent frameworks, and the economics of running models at scale.
HN discussions here tend to be unusually practical — benchmark comparisons, cost breakdowns, and real deployment war stories from teams shipping LLM features.
158 stories archived · Page 1 of 6
RSS feed for LLMVyne – Zapier for DeFi users, run from any LLM or the workflow builder
app.vyne.finance · 2 points · 0 comments
I sent Claude Opus 5 '–-' and it wrote me 5k tokens about a cartographer
austinsnerdythings.com · 15 points · 5 comments
OpenReviewer: A Specialized LLM for Generating Critical Scientific Paper Reviews
aclanthology.org · 4 points · 0 comments
What's Wrong with American Studies?
radicallypragmatic.org · 5 points · 1 comments
What if useful AI is a fantasy?
lzon.ca · 9 points · 3 comments
Writekin – fine-tune a local LLM on your own writing, on your Mac
github.com · 2 points · 0 comments
"Uncensored" open LLMs are measurably more optimistic than their base models
arxiv.org · 23 points · 7 comments
SOTA on the hardest AI memory benchmark (BEAM, 10M tokens), with a smaller model
2 points · 3 comments
Now Is the Time to Give LLMs Access to the ACM Digital Library
cacm.acm.org · 85 points · 62 comments
Ctrlb-decompose: Strip the noise from logs before sending to LLMs
github.com · 48 points · 6 comments
Claude Code system prompt says not to ask for permission, assumes user is absent
elliotmilco.substack.com · 4 points · 0 comments
Mondragon Corporation – a federation of co-operatives
en.wikipedia.org · 161 points · 31 comments
DMARC has been public since 2012 but most company domains still don't enforce it
ciphercue.com · 160 points · 100 comments
Does Speaking to Agents Like Cavemen Save 65% of Tokens? We Test
blog.jetbrains.com · 5 points · 1 comments
Orchard – Let AI agents set up your app's back end with one prompt
orchard-dashboard.pages.dev · 2 points · 0 comments
Day 0 Kimi-K3 Inference Deployment with Atom on AMD Instinct MI355X GPUs
amd.com · 9 points · 0 comments
PromptTrace – Free hands-on labs to practice hacking LLMs
prompttrace.airedlab.com · 3 points · 0 comments
Don't ask an LLM for a confidence score
justinflick.com · 20 points · 1 comments
Measured LLM inference speeds on Apple Silicon, with raw data (CC BY 4.0)
macyou.co · 10 points · 3 comments
Kimi K3 Now Available via Telnyx Inference API
telnyx.com · 69 points · 24 comments
Professor's invisible prompt trap catches 32/35 students cheating with AI
techspot.com · 81 points · 74 comments
Evading Residential Proxy Networks
fbi.gov · 4 points · 2 comments
Kimi K3 on vLLM: Up to 370 Tokens/sec
vllm.ai · 5 points · 0 comments
Ctxdiff – Git diff for your LLM agent's context window
github.com · 3 points · 3 comments
Where do those LLM-generated outreach emails come from?
3 points · 1 comments
General Resolution: LLM Usage in Debian
debian.org · 3 points · 0 comments
ASD-STE100 Simplified Technical English for LLMs
github.com · 5 points · 2 comments
Are we having a substantial increase of "Show HN" since LLM aided coding
7 points · 4 comments
Sand battery: Finland's answer to a renewable energy headache
cnbc.com · 27 points · 8 comments
Aqua UI for the Web
github.com · 2 points · 0 comments