Editorial desk

GenAI Brief

The GenAI Brief editorial desk reads the launches, papers and pricing pages so you do not have to, and writes down what actually changed.

  1. NewsAnalysis

    Inference prices keep falling. Here is who actually benefits

    Per-token prices for frontier-class models have dropped by roughly an order of magnitude every 18 months. The savings are real, but they land unevenly across the stack.

    4 min read

  2. NewsExplainer

    The EU AI Act's model rules, one year in: what changed for builders

    General-purpose AI model obligations under the EU AI Act have applied since August 2025. A year later, here is what providers have actually had to do, and what still applies to you if you only fine-tune or deploy.

    3 min read

  3. ModelsExplainer

    How to read a model card without getting fooled

    Model cards are part specification, part marketing. Here is a field-by-field guide to the numbers that matter, the ones that are routinely gamed, and the questions a card should answer before you ship on it.

    4 min read

  4. ModelsAnalysis

    Open-weight models are closing the gap. The economics say it will not fully close

    The lag between the best closed model and the best open-weight model has shrunk to months on most benchmarks. Whether it reaches zero depends on who pays for frontier training runs and why.

    3 min read

  5. ResearchExplainer

    What "reasoning" models actually do differently

    Reasoning models are trained to spend tokens thinking before they answer. Here is what that training involves, why it works on some problems and not others, and how to decide when to pay for it.

    4 min read

  6. ResearchAnalysis

    Test-time compute changed the scaling roadmap. Here is what it costs

    For a decade, progress meant bigger training runs. Now labs can trade inference compute for capability instead. That shifts the economics from capex at the lab to opex at the user, and it changes what "a better model" means.

    3 min read

  7. ToolsComparison

    Five LLM gateways compared: routing, failover, governance and where each fits

    LLM gateways sit between your applications and model providers to handle routing, keys, failover, budgets and logging. We compare Bifrost, LiteLLM, Portkey, Kong AI Gateway and Cloudflare AI Gateway on the decisions that actually differ.

    4 min read

  8. ToolsSurvey

    The state of AI agent frameworks in 2026: a survey

    Agent frameworks have split into graph-based orchestrators, lab-native SDKs, multi-agent role systems and typed minimalists. This survey maps the landscape, the design bets behind each camp and the questions to ask before committing.

    4 min read

  9. IndustryEditorial

    Benchmarks became a marketing channel. Treat them like one

    Benchmark tables were meant to be measurements. They are now launch collateral, optimised for by every lab and reported under whatever settings look best. That does not make them useless, but it changes how they should be read.

    3 min read

  10. IndustryExplainer

    The real cost of running an AI product, line by line

    Token spend is the line everyone watches and rarely the largest. A working breakdown of where the money goes in a production generative AI product, from inference and evaluation to the humans in the loop.

    3 min read