Reinvently
News Guides LLM Leaderboard About Subscribe

Reinvently

ReinventlyReinventlyReinvently

Reinvently

ReinventlyReinventlyReinvently

Putting the Intelligence into AI

Putting the Intelligence into AIPutting the Intelligence into AIPutting the Intelligence into AI

Putting the Intelligence into AI

Putting the Intelligence into AIPutting the Intelligence into AIPutting the Intelligence into AI

Independent research

Latest Analysis

Original benchmarks, tested tools and practical guidance on AI engineering, UK governance and enterprise adoption.

16 July 2026  ·  Model Evaluation

Building the Ed-o-meter: Notes on Writing My Own LLM Benchmark

The harness, the tasks, the decisions that turned out to matter, and the mistakes — written down mostly because the mistakes were more informative than the results.

Read more

5 July 2026  ·  Technology Strategy

Locked-Down Fable Disappoints, Cheap-as-Chips GLM-5.2 Unsafe, GPT-5.5 Wins as the All-Rounder

GPT-5.5 passed 28 of 28, GLM-5.2 passed 26 at an eighth of the cost, and Fable 5 refused five perfectly benign tasks

Read more

2 July 2026  ·  UK AI News

The UK AI Policy Landscape: What Enterprise Leaders Need to Track in 2026

There is no UK AI Act, and none is coming. What exists instead is an estate — regulators, compute money, sandboxes, institutes — that only makes sense viewed whole.

Read more

2 July 2026  ·  UK AI News

AI Growth Labs: What the UK's Regulatory Sandboxes Mean for Your Sector

Regulatory sandboxes for AI open to legal services this summer, with other sectors to follow. What direct access to your regulator buys, and what to do before applications open.

Read more

2 July 2026  ·  Enterprise Adoption

The Agent Governance Gap: $202m Budgets, 26 Percent Cost Visibility

Q2 2026 data shows agent deployment plateauing, orchestration doubling and employee resistance quadrupling — with the money committed before the controls exist.

Read more

2 July 2026  ·  Technology Strategy

Is Claude Fable 5 Worth It for Enterprise Coding?

Fable 5 leads every coding benchmark here by a wide margin, at double the price of Anthropic's own Opus 4.8. Whether the premium is worth it, and where GLM-5.2 does the job just as well.

Read more

1 July 2026  ·  Technology Strategy

How to Sandbox AI Agent Code: Firecracker, OpenSandbox, Docker, SmolVM and nono Compared

The brief said microVMs, but only two of the five are. A decision guide to the four isolation models for running AI agent code — and how to pick the right one.

Read more

6 June 2026  ·  Technology Strategy

Multi-Agent Orchestration Frameworks Compared: ruflo, Aperant, Sandcastle, Mission Control and Maestro

An evidence-based comparison of five multi-agent orchestration tools — ruflo, Aperant, Sandcastle, Mission Control and Maestro — by popularity, use case, hosting and community feedback.

Read more

30 May 2026  ·  Technology Strategy

GSD, BMAD, OpenSpec, or GitHub Spec Kit: Choosing the Right AI Development Framework

Four spec-driven AI development frameworks have emerged as the serious options for structured AI coding. They share a starting principle — and diverge sharply on everything else.

Read more

21 May 2026  ·  Enterprise · Technology Strategy

Claude Memory and Dream: Evidence, Architecture and Risk

Reported deployment gains, system architecture, and the governance risks of persistent, self-editing agents.

Read more

15 May 2026  ·  UK AI · Enterprise

UK AI Regulation Tightens: What the ICO Guidance and Parliament's Inquiry Mean for Employers

The ICO, Parliament, and the DRCF are reshaping what UK employers must do when AI touches decisions that affect people.

Read more

8 May 2026  ·  Enterprise

The Enterprise AI Reality Check: Why 79% of Organisations Aren't Seeing ROI

AI funding hit $300 billion in Q1 2026. Meanwhile, 79% of organisations face significant adoption challenges. Here is what the successful 29% are doing differently.

Read more

29 April 2026  ·  UK AI · Enterprise

UK Sovereign AI: What the Government's £500m Bet Means for Your Organisation

Liz Kendall's £500m Sovereign AI programme frames compute concentration as a national security risk. Here is what it signals for enterprise technology strategy.

Read more

24 April 2026  ·  Technology Strategy

GPT-5.5 and the Age of Autonomous Task Completion

OpenAI's GPT-5.5 is the first mainstream model designed for AI that acts rather than assists. Here is what the governance and workflow implications mean for enterprise teams.

Read more

22 April 2026  ·  Enterprise · Technology Strategy

A2A Protocol: An Enterprise Architecture Decision Guide

What A2A standardises, where interoperability still stops, and how to evaluate vendor support in architecture and procurement.

Read more

20 April 2026  ·  AI · Global

Global AI Adoption: How Different Countries Are Embracing Artificial Intelligence

From US velocity to EU caution, Chinese state strategy to Middle Eastern sovereign ambition — a country-by-country breakdown.

Read more

19 April 2026  ·  AI · Cybersecurity

Claude Mythos and Project Glasswing: What the Security Evidence Shows

Reported vulnerability findings, limits of the available data, and implications for security teams.

Read more

13 April 2026  ·  AI · UK

UK AI Companies: A Field Guide to Ten Organisations

An unranked selection used to examine what the UK ecosystem produces and where it remains constrained.

Read more

12 April 2026  ·  AI · Software Engineering

RAG vs GraphRAG: Which Retrieval Architecture Is Right for Your AI Application?

Retrieval-augmented generation transformed enterprise AI. Now GraphRAG is challenging the assumptions it was built on.

Read more

5 April 2026  ·  AI · Enterprise

Copilot Studio vs Azure AI Foundry vs AWS Bedrock

A practical guide to choosing the right enterprise AI platform for your organisation.

Read more

14 March 2026  ·  AI · Education

The UK's AI Skills Gap: The Biggest Opportunity of the Decade

73% of UK adults have had no AI training at all. Here's where the opportunity lies — and where to start.

Read more

27 May 2026  ·  AI · Software Engineering

GitHub Copilot vs Claude Code vs Cursor

A practical comparison of three AI coding tools — with real-world pricing models and costs at scale.

Read more

10 February 2026  ·  AI · Policy · UK

The UK's Public AI Infrastructure: Government Institutes Shaping the National AI Ecosystem

From the Alan Turing Institute to the AI Security Institute — what the UK's publicly funded AI bodies actually do and why they matter.

Read more

18 February 2026  ·  AI · Enterprise · Microsoft

Microsoft Foundry: Architecture, Governance and Trade-offs

A decision guide to Foundry's control plane, model and agent services, Azure advantages, operational costs and trade-offs.

Read more

22 January 2026  ·  AI · Enterprise

Generative AI Adoption by Industry

A sector-by-sector breakdown covering financial services, healthcare, retail, legal, charities and more.

Read more

12 November 2025  ·  AI

AI-Assisted Web Development: Where It Helps and Where It Fails

Where coding and content tools save time, where they create risk, and which review controls matter.

Read more
View all posts

Get New Research by Email

Get new benchmarks, UK AI governance analysis and engineering guides when they are published.

Subscribe →