All topics

LLM stories

Large language models have moved from research curiosity to production infrastructure. This topic focuses on LLM APIs, fine-tuning, RAG pipelines, agent frameworks, and the economics of running models at scale.

HN discussions here tend to be unusually practical — benchmark comparisons, cost breakdowns, and real deployment war stories from teams shipping LLM features.

898 stories archived · Page 28 of 30

RSS feed for LLM

Promptrack A local menu bar app that tracks your Claude Code usage

promptrack.dev · 3 points · 0 comments

Autograd-Free LLM Guiding with 0MB VRAM (Alternative Pathways)

github.com · 2 points · 1 comments

TTFT benchmark: LLM Gateway vs. OpenRouter (Claude-haiku-4.5, 150 runs)

llmgateway.io · 3 points · 0 comments

Millwright – Rust-based, self-hosted LLM router

github.com · 10 points · 7 comments

Chrome Extension Claude Token Usage Bar and Context Use for Claude.ai

github.com · 2 points · 0 comments

GigaToken: ~1000x faster Language model tokenization

github.com · 619 points · 119 comments

Can a MUD evaluate LLMs? A $99 proof of concept

cruciblebench.ai · 109 points · 79 comments

Ghost Cut – Or why Cut and Paste is broken everywhere

ishmael.textualize.io · 196 points · 154 comments

Lucen a Python compiler that parallelizes for-loops via comment pragmas

github.com · 10 points · 17 comments

Controlling Reasoning Effort in LLMs

magazine.sebastianraschka.com · 84 points · 8 comments

Altman: GPT-5.6 is 54% more token efficient on agentic coding

cnbc.com · 11 points · 4 comments

TensorRT-LLM running natively on Windows (no WSL)

baremetalrt.ai · 2 points · 0 comments

ContextNest versioned, governed context for AI agents (open-source CLI)

promptowl.ai · 3 points · 0 comments

Tokenstead, find AI models for your hardware

tokenstead.ai · 3 points · 2 comments

LLMs for technical editing: The good, the bad, and the ugly

techstackups.com · 3 points · 0 comments

Slopera, a browser that hallucinates every page with an LLM

github.com · 3 points · 1 comments

OpenTab – a lazygit-style TUI for your AI token spend

github.com · 2 points · 0 comments

Battle LLM Robots – Prompt your LLM, Submit your bot, Watch it battle

battlellmrobots.com · 4 points · 0 comments

A control-theory approach to detecting LLM agent instability

github.com · 2 points · 0 comments

SoulOS–State&persona management for agents without giving up your LLM

mziqudhd92.github.io · 2 points · 2 comments

AI software that generates 'rage bait' developed by Germany's far-right AfD

irishtimes.com · 5 points · 0 comments

I Think I Have LLM Burnout

alecscollon.com · 9 points · 363 comments

Alphaville LLC initiates coverage of SpaceX with Buy recommendation

ft.com · 5 points · 2 comments

Agentic test processes, LLM benchmarks, and other notes on agentic coding fr

danluu.com · 10 points · 2 comments

Onboard-CLI, a LLM powered and AST-based tool to visualize codebase

github.com · 11 points · 14 comments

Ramp.com Website Is a Prompt

web.archive.org · 5 points · 3 comments

How you manage local long lived research projects and LLM's?

5 points · 2 comments

Foreman, a self-hosted LLM gateway for cost aware model routing

github.com · 5 points · 16 comments

Germans turn to battery storage to shield against fossil fuel price shocks

euronews.com · 3 points · 0 comments

Reform UK leader 'in real trouble' against Count Binface

mirror.co.uk · 32 points · 36 comments