LLM stories
Large language models have moved from research curiosity to production infrastructure. This topic focuses on LLM APIs, fine-tuning, RAG pipelines, agent frameworks, and the economics of running models at scale.
HN discussions here tend to be unusually practical — benchmark comparisons, cost breakdowns, and real deployment war stories from teams shipping LLM features.
158 stories archived · Page 5 of 6
RSS feed for LLMSoulOS–State&persona management for agents without giving up your LLM
mziqudhd92.github.io · 2 points · 2 comments
AI software that generates 'rage bait' developed by Germany's far-right AfD
irishtimes.com · 5 points · 0 comments
I Think I Have LLM Burnout
alecscollon.com · 9 points · 0 comments
Alphaville LLC initiates coverage of SpaceX with Buy recommendation
ft.com · 5 points · 2 comments
Agentic test processes, LLM benchmarks, and other notes on agentic coding fr
danluu.com · 10 points · 1 comments
Onboard-CLI, a LLM powered and AST-based tool to visualize codebase
github.com · 11 points · 2 comments
Ramp.com Website Is a Prompt
web.archive.org · 3 points · 2 comments
How you manage local long lived research projects and LLM's?
5 points · 1 comments
Foreman, a self-hosted LLM gateway for cost aware model routing
github.com · 5 points · 2 comments
Germans turn to battery storage to shield against fossil fuel price shocks
euronews.com · 3 points · 0 comments
Reform UK leader 'in real trouble' against Count Binface
mirror.co.uk · 32 points · 19 comments
Chorus: A fast WAL for object storage
rockwotj.com · 5 points · 4 comments
Instant GraphRAG over any Postgres database
polygres.com · 5 points · 4 comments
Men's average testosterone levels have halved in last 50 years
theguardian.com · 66 points · 68 comments
LOL Storage Bug on Microsoft Windows 11 Could Eat Up 500 GB Disk Space
itsfoss.com · 12 points · 1 comments
Bike4Mind – open-core AI workbench; any model, agents, RAG, self-host
github.com · 3 points · 2 comments
What is your AI harness that lets you switch LLM models easily?
10 points · 8 comments
PL/Ruby: Ruby as a procedural language for PostgreSQL (functions, triggers, SPI)
github.com · 6 points · 0 comments
I wrote a 1-bit WebGPU runtime to run a 1.7B LLM in the browser
aidekin.com · 5 points · 2 comments
Baerly-storage, a document DB that runs per request, no DB server
github.com · 16 points · 1 comments
Tracking GenAI cost and endpoint fragility so app teams don't have to
llmintel.ai · 2 points · 0 comments
Are LLMs slowly making companies dysfunctional?
7 points · 3 comments
I built a free website that makes LLM prompting easier in 40 languages
enlive.inc · 3 points · 1 comments
Signal for LLM – The Modulator Architecture (Theory Complete)
divinecanon.github.io · 2 points · 0 comments
Frugon – Find which LLM calls a cheaper model could handle (local, MIT)
github.com · 7 points · 2 comments
Reinforcement Learning with Metacognitive Feedback Elicits Uncertainty in LLMs
arxiv.org · 12 points · 1 comments
Shadow Web – Cut 64–97% of web page tokens for LLM agents
github.com · 2 points · 0 comments
A little cat that counts your tokens (Claude and codex)
jpthecat.com · 3 points · 0 comments
Onboard CLI uses LLM to filter out nodes and AST to visualize codebase
github.com · 2 points · 0 comments
RagPack – Lightweight self-hosted RAG infra for startups
github.com · 2 points · 0 comments