Editorial desk
GenAI Brief
The GenAI Brief editorial desk reads the launches, papers and pricing pages so you do not have to, and writes down what actually changed.
NewsAnalysis
Inference prices keep falling. Here is who actually benefits
Per-token prices for frontier-class models have dropped by roughly an order of magnitude every 18 months. The savings are real, but they land unevenly across the stack.
NewsExplainer
The EU AI Act's model rules, one year in: what changed for builders
General-purpose AI model obligations under the EU AI Act have applied since August 2025. A year later, here is what providers have actually had to do, and what still applies to you if you only fine-tune or deploy.
ModelsExplainer
How to read a model card without getting fooled
Model cards are part specification, part marketing. Here is a field-by-field guide to the numbers that matter, the ones that are routinely gamed, and the questions a card should answer before you ship on it.
ModelsAnalysis
Open-weight models are closing the gap. The economics say it will not fully close
The lag between the best closed model and the best open-weight model has shrunk to months on most benchmarks. Whether it reaches zero depends on who pays for frontier training runs and why.
ResearchExplainer
What "reasoning" models actually do differently
Reasoning models are trained to spend tokens thinking before they answer. Here is what that training involves, why it works on some problems and not others, and how to decide when to pay for it.
ResearchAnalysis
Test-time compute changed the scaling roadmap. Here is what it costs
For a decade, progress meant bigger training runs. Now labs can trade inference compute for capability instead. That shifts the economics from capex at the lab to opex at the user, and it changes what "a better model" means.
ToolsComparison
Five LLM gateways compared: routing, failover, governance and where each fits
LLM gateways sit between your applications and model providers to handle routing, keys, failover, budgets and logging. We compare Bifrost, LiteLLM, Portkey, Kong AI Gateway and Cloudflare AI Gateway on the decisions that actually differ.
ToolsSurvey
The state of AI agent frameworks in 2026: a survey
Agent frameworks have split into graph-based orchestrators, lab-native SDKs, multi-agent role systems and typed minimalists. This survey maps the landscape, the design bets behind each camp and the questions to ask before committing.
IndustryEditorial
Benchmarks became a marketing channel. Treat them like one
Benchmark tables were meant to be measurements. They are now launch collateral, optimised for by every lab and reported under whatever settings look best. That does not make them useless, but it changes how they should be read.
IndustryExplainer
The real cost of running an AI product, line by line
Token spend is the line everyone watches and rarely the largest. A working breakdown of where the money goes in a production generative AI product, from inference and evaluation to the humans in the loop.









