LLM stories
Large language models have moved from research curiosity to production infrastructure. This topic focuses on LLM APIs, fine-tuning, RAG pipelines, agent frameworks, and the economics of running models at scale.
HN discussions here tend to be unusually practical — benchmark comparisons, cost breakdowns, and real deployment war stories from teams shipping LLM features.
898 stories archived · Page 28 of 30
RSS feed for LLMPromptrack A local menu bar app that tracks your Claude Code usage
promptrack.dev · 3 points · 0 comments
Autograd-Free LLM Guiding with 0MB VRAM (Alternative Pathways)
github.com · 2 points · 1 comments
TTFT benchmark: LLM Gateway vs. OpenRouter (Claude-haiku-4.5, 150 runs)
llmgateway.io · 3 points · 0 comments
Millwright – Rust-based, self-hosted LLM router
github.com · 10 points · 7 comments
Chrome Extension Claude Token Usage Bar and Context Use for Claude.ai
github.com · 2 points · 0 comments
GigaToken: ~1000x faster Language model tokenization
github.com · 619 points · 119 comments
Can a MUD evaluate LLMs? A $99 proof of concept
cruciblebench.ai · 109 points · 79 comments
Ghost Cut – Or why Cut and Paste is broken everywhere
ishmael.textualize.io · 196 points · 154 comments
Lucen a Python compiler that parallelizes for-loops via comment pragmas
github.com · 10 points · 17 comments
Controlling Reasoning Effort in LLMs
magazine.sebastianraschka.com · 84 points · 8 comments
Altman: GPT-5.6 is 54% more token efficient on agentic coding
cnbc.com · 11 points · 4 comments
TensorRT-LLM running natively on Windows (no WSL)
baremetalrt.ai · 2 points · 0 comments
ContextNest versioned, governed context for AI agents (open-source CLI)
promptowl.ai · 3 points · 0 comments
Tokenstead, find AI models for your hardware
tokenstead.ai · 3 points · 2 comments
LLMs for technical editing: The good, the bad, and the ugly
techstackups.com · 3 points · 0 comments
Slopera, a browser that hallucinates every page with an LLM
github.com · 3 points · 1 comments
OpenTab – a lazygit-style TUI for your AI token spend
github.com · 2 points · 0 comments
Battle LLM Robots – Prompt your LLM, Submit your bot, Watch it battle
battlellmrobots.com · 4 points · 0 comments
A control-theory approach to detecting LLM agent instability
github.com · 2 points · 0 comments
SoulOS–State&persona management for agents without giving up your LLM
mziqudhd92.github.io · 2 points · 2 comments
AI software that generates 'rage bait' developed by Germany's far-right AfD
irishtimes.com · 5 points · 0 comments
I Think I Have LLM Burnout
alecscollon.com · 9 points · 363 comments
Alphaville LLC initiates coverage of SpaceX with Buy recommendation
ft.com · 5 points · 2 comments
Agentic test processes, LLM benchmarks, and other notes on agentic coding fr
danluu.com · 10 points · 2 comments
Onboard-CLI, a LLM powered and AST-based tool to visualize codebase
github.com · 11 points · 14 comments
Ramp.com Website Is a Prompt
web.archive.org · 5 points · 3 comments
How you manage local long lived research projects and LLM's?
5 points · 2 comments
Foreman, a self-hosted LLM gateway for cost aware model routing
github.com · 5 points · 16 comments
Germans turn to battery storage to shield against fossil fuel price shocks
euronews.com · 3 points · 0 comments
Reform UK leader 'in real trouble' against Count Binface
mirror.co.uk · 32 points · 36 comments