These startups are chasing the next big thing in LLMs

MIT Technology Review – AI(Global) 10 Aug 2026 38

Emerging LLM architectures promising large efficiency gains could reshape the cost and capability assumptions underpinning agency AI business cases.

  • Diffusion LLMs generate whole text blocks simultaneously, claiming 10x speed and cost gains over standard transformers.
  • Google's Diffusion Gemma prototype suggests major labs are validating this architectural direction alongside startups.
  • Limited direct governance or procurement implications for APS agencies at this stage - primarily a technology watch item.
  • Monitor Technology and procurement teams may want to monitor whether diffusion LLM efficiency claims mature into commercially available models that could affect agency AI cost modelling.

Implications are AI-generated. Starting points, not advice — see methodology for how they're framed.

View original source