The Token Bill: What Your LLM App Really Costs (and How to Measure It)
Token pricing looks simple until you're running a million calls a month. Here's a practical way to count the real cost — and a small harness for trading cost against quality.
The journal
Essays with runnable code and interactive demos — from React frontends to JVM backends, data pipelines to LLM-powered features. Written by a fullstack developer working in AI. Browse by category or tag.
Token pricing looks simple until you're running a million calls a month. Here's a practical way to count the real cost — and a small harness for trading cost against quality.
Your code runs top to bottom. Spark doesn't. A small-dataset walkthrough of lazy execution order, shuffles, caching, persist, checkpointing — and the best practices that fall out of them.
JDK 27 shipped September 15, 2026 with 9 JEPs. Four are final — and two of them change how every JVM you run behaves without touching a line of code.
NVIDIA's NOOA framework treats an AI agent as a plain Python object: fields are state, methods are capabilities, docstrings are prompts, and an ellipsis body means 'the LLM implements this.' Here's how it works and how to build your first agent.