AIAI EngineerMay 18, 2026· 28:18

Rewiring the State — Eoin Mulgrew, No. 10 (Downing Street)

Eoin Mulgrew, from the Number 10 Data Science team, details how a small insurgent unit at the center of UK government bypasses bureaucracy to rapidly deploy AI. The team recruits exclusively outsiders (0.7% acceptance rate), pays market rates, and ships tools in weeks. Examples include an engineer who saved £1.5M by building a statute-book analysis tool in two weeks, a policy simulation platform for Universal Credit, a delivery red-teaming PMO, and a public service that went from idea to live in two months. The team also placed fellows in the AI Safety Institute, the Incubator for AI (co-creating Xtract with DeepMind to digitize planning applications), and Justice AI, which embeds engineers in prisons. Mulgrew closes with Will, a Y Combinator founder and Harvard dropout, standing outside HMP Wormwood Scrubs with the keys two weeks into the job, urging talented technologists to 'join us, and we'll give you the keys to the state.'

  1. 0:00Intro
  2. 1:58The Challenge
  3. 4:58Insurgency
  4. 8:43Our Approach
  5. 9:56Quick Wins
  6. 14:42Partnerships
  7. 18:12Justice AI
  8. 20:11Join Us
  9. 21:10Q&A

Powered by PodHood

Transcript

Intro0:00

Eoin Mulgrew0:16

Good to go. Can people hear me okay? Awesome. Cool, thank you. I'm slightly embarrassed by the grandiose title; I know that I've had to leave it up for a few seconds. But anyway, who here works in government, can I ask?

Show of hands. Very good. That's what I was hoping for. You might be very well acquainted with some of the stuff that I'm going to grumble about. Who can't think of anything worse?

Alright. Good. What gets measured gets improved, so I'll do another one at the end. Actually, I won't, in case more hands go up. Hi everyone, I'm Eoin Mulgrew. I work in the data science team just down the road in 10 Downing Street.

I run our cross-government transformation work, including our fellowship program, which is predominantly what I'm going to talk about today. Yeah, I wanted to come down here and tell you what we're doing, in the hope that some of you might decide you want to be a part of it, which would be quite cool.

Quick bit about us. So, Number 10 Data Science Team, 10DS. We were set up sort of during the pandemic, partly in a response to the pandemic. Our core business is making sure that the most important decisions in the country are informed by the best possible evidence.

However, we are in the process of quite radically scaling up our own AI engineering and development capability, not just with the intention of driving AI adoption within Number 10 itself, but also across strategically important parts of the state.

And the way that we're going about doing that is quite novel in itself. Before we get into that, just a little bit of context to set the scene. I know some of you are flying in from the West Coast, etc.

Some of you might not follow the news. Believe it or not, there are some challenges when it comes to public service delivery in the UK.

The Challenge1:58

Eoin Mulgrew2:08

I say that half in jest, but it's pretty serious. You know, at the moment there are 7.25 million people that are on NHS waiting lists. There are, I think, about 350,000 court cases that are stuck in a backlog.

Only 1 in 5 planning application decisions in this country are currently decided on time. Sitting behind all of this is a public sector productivity crisis that was bad and has only been exacerbated since the pandemic.

There are different figures for the sort of extent of this crisis. I've gone for a Tony Blair Institute figure that says there's a sort of $40 billion prize in annual productivity gains from AI in government. But it's clear to anybody that works in the system and most of society that if any industry is ripe for disruption over the next few years, government is one.

And I call it an industry rather than an organization. It's a big, complex industry of 400,000 people, and I think we should look at it through that lens. Unfortunately, however, government has traditionally not been great at building and nurturing high-performing technical teams.

A lot of these issues are not specific to the UK. A lot of our American friends will be familiar with them, but just some of the commonly cited ones. Pay is an obvious one. It makes it quite hard for us to compete for the best talent out there and then retain it.

But also some barriers that are both real and perceived. So government is a very hierarchical organization in many parts. There's a lot of bureaucracy. As a result, it can often move incredibly slowly, not just for those reasons, but also the fact that there are regulations and safeguards in place that are very sensible because we're ultimately accountable to the public and to Parliament.

But all of this can result in a system that is not always that appetizing for high-performing technical people to join, especially the sort of people that we want, people that are impatient to leave their mark on the world.

So what do we do with that? Or what do we do about that?

Changing a lot of these things, it's quite like a systemic challenge. It's like turning an oil tanker, which is a bit of a tired cliché, but it's a good one. We're a little team at the center. Turning that oil tanker is beyond the remit of any one team, let alone a scrappy little startup like ours.

However, I think Cal Beer, our Chief AI Officer, is going to be closing out the conference this evening. That's very much his job, and he's doing great work at the Department of Science and Technology to do just that.

So I recommend everybody goes along to it. But there is a lot of political will at the moment to get stuff done and to make sure that this time around we are seizing new technology to actually make a dent in some of those problems I just mentioned.

Insurgency4:58

Eoin Mulgrew5:13

So the question put to us was, well, what can you do about it? And this was sort of the answer.

As I said, we're like quite a small team at the center. In terms of what we can do, we said, okay, well, let us take the shackles off. Let us basically set up a small insurgent unit at the very center that is not sort of burdened by some of the constraints that I just described to you.

What do I mean by an insurgency model? So we're setting up a new team. It operates with a mandate from Number 10. We operate with an unusually high level of political backing to go into departments and get stuff done.

We're able to pay market rates within reason. We're not paying like meta money necessarily. But the thing is, a lot of people will happily take pay cuts if we make it economically viable to come in and work on some of these challenges because they're interested,right?

We operate with an unusually high level of autonomy as well. We're able to be fairly opportunistic about the challenges we take on, where we go into a department and see opportunities to have impact. Also, this one's pretty crucial.

The civil service standard recruitment process is optimized for a lot of things, but not necessarily recruiting exceptional technical talent. We've been allowed to recruit our own way. We've got a fairly grueling selection process that is laser-targeted on technical skills.

We've got a success rate of about 0.7, 0.8 percent. And most interestingly of all, and this is what differentiates us, we recruit exclusively outsiders. One of the best ways that I can have impact is by getting some people of the likes of this conference into government, because what has happened in the past is they tend not to leave, and some of them end up setting up their own teams.

I'll get into that later. Just when we're on this, though, I don't want to make this sound like overly simplistic. A lot of people from the outside, particularly the tech industry, think that this bit alone is the only important thing that you need, that if you have a big enough stick for ministers, you can go in, you can break down data silos, you can do what you want.

In practice, it's a lot harder than that. Otherwise, everybody would be doing it.

And it's really early days. Like, we're only setting out on this journey, but it turns out there's a huge amount of appetite for it. So we have been taking people from the labs. We've been taking people from big tech, from top research institutes.

We've been taking YC founders, serial entrepreneurs, people who probably did not think they would be working in the civil service this time last year. But when you think about it, the decisions that go across a minister's desk are like some of the most important things you could possibly work on.

So if you make it economically viable and you promise people that you're going to put them in an environment where they can do their best work, it makes it really interesting. It's also worth pointing out as well, we do want to recruit missionaries, not mercenaries.

So the pay matters, but it's not alone, because a paycheck is not going to get you out of bed in the morning when stuff gets hard. And doing the stuff that we do does tend to be difficult. In terms of how we operate then, again, a bit different from normal government teams.

Our Approach8:43

Eoin Mulgrew8:43

Some of the, there's like an abundance of low-hanging fruit around the system. As you can imagine, it's a legacy organization. There's lots of simple AI use cases that you can do in a few days to save money, to improve service delivery, all that good stuff.

When it comes to that, we largely do it ourselves. That's the easiest and most satisfying part. As we speak, we've actually got the first forward-deployed engineers in the history of 10 Downing Street embedding themselves with policy and operational teams, teams of policy advisors, teams of lawyers, teams of comms people, pollsters, and everything in between.

They're observing their workflows, their pain points, co-designing solutions with them to help them do their job more efficiently and effectively, and generally taking things from idea to implementation in a couple of weeks and getting new capability into the hands of users quickly.

And then some of the other problems we talked about. I mentioned some of the huge backlogs in the system. That is not low-hanging fruit. That's really complicated stuff. And normally, when it comes to those, we take more of a partnership model where we will deploy some of our people into another team or another department, sometimes for prolonged periods.

Quick Wins9:56

Eoin Mulgrew9:56

And I'm going to give you examples of both of these in a minute. I'm going to start with some of the low-hanging fruit. It's worth pointing out, actually, a lot of the stuff that we do in Number 10, I can't really show you.

I know that sounds like an easy get-out-of-jail card, but trust me, some of the stuff is a little bit sensitive. We're doing a lot of workflow automation, augmenting existing teams, as you can imagine. And here are a few other examples of stuff we've done just in the past few weeks.

So policy simulation, that's turned out to be really interesting. So here we can allow policy teams in the building to test out the impact of different policy decisions before they're made. I think in this one here, we're looking at different decisions around universal credit and how they might impact, oh, I've paused it.

How they might impact household finances, amongst other things. But this can be applied to a broad range of stuff. Not replacing human analysis necessarily. We're not putting ourselves out of a job. But what it has meant is that far more decisions in the building are being informed by high-quality modeling and at a far faster rate than otherwise would have been the case.

And this is another one just from the past couple of weeks. So the Cabinet Office was about to spend 1.5 million pounds on getting an outside firm of lawyers to come in and do analysis of the entire UK statute book.

Granted, the statute book is the height of four African elephants of legalese, but still a pretty obvious AI use case. So we were going to spend 1.5 million. Instead, one of our engineers embedded with that team of in-house lawyers for a couple of weeks.

The benefits of this is not just money saved. Obviously, 1.5 million is not nothing, but also speed. So the issue with the analysis that we were going to pay for is that it was going to be done slower than the pace at which new laws and regulations are made, which means you're going to have to do it again after a certain period of time.

So now we've got this tool that that team can use, and they can do it whenever they want at the drop of a hat. And we can also potentially open source it and share it with other teams in government.

And then this is another one. So in Number 10, we're responsible for the delivery of every major project and manifesto commitment in government. That means a lot of reports come in on how various things are doing. This is a little sort of delivery red-teaming tool that the team spun up a couple of weeks ago that is now being used every day.

It's essentially a PMO that we've put in the pockets of delivery teams in Number 10, not just so that they can interrogate the delivery reports that are coming across their table, but also give a second judgment on the teams that are reporting them.

So it will flag up to decision makers in Number 10, does this team, does this department normally have a bit of optimism bias? Do they tend to disproportionately rate their risks as amber? And are their mitigations usually effective or not?

Also, aside from AI adoption, having this capability in-house is really good. I think transparency is one thing that this country can do a bit better at. Up until a couple of months ago, the government had never published a public-facing dashboard so that you lot can actually see how we're doing when it comes to delivery.

But now we've published two in as many months. I think some of you might be familiar with the one on the left. This is the AI Opportunities Action Plan that Matt Clifford drafted about a year ago. This is how the UK is doing when it comes to rolling out compute and generally setting up the UK to be a leader in AI adoption.

And now you can go online and see how we're actually doing. Yeah. Also, another thing that I can't show, but in two and a half weeks' time, one of our ministers is going to launch a new public service that millions of people in the country are going to use.

I can't go into more detail and steal their thunder. But in their words, it's hard to believe this didn't already exist. I assumed something like it already existed. That's something that we thought of two months ago and is now going to be live and used by the public.

It is not an understatement to say that normally in government, a project like that might be in discovery for a year or more. So yeah. Let's get into the meatier stuff then. So that's nice low-hanging fruit within the building.

Now I want to talk about some of the work that we're helping other teams with across different parts of the ecosystem. For the purposes of this, I'm just going to focus on three of our partners, the AI Safety Institute, the Incubator for AI, and Justice AI.

Partnerships14:42

Eoin Mulgrew15:01

YC, I think pretty much everybody in this room will be familiar with it. Massive win for the UK. It's a great thing that we set up. We're a leading government body for evaluating frontier models, and it was also the world's first.

And we were really proud from day one to support it by putting a couple of our fellows in there to help them set up their cybersecurity workstream, amongst other things. I'll not dwell too much on this, but one of our early fellows was Dr.

Harry Koppock. I don't think Harry's here, but we put him into YC from day one. And he led on their inspect tool, amongst other things. So there's a safe, isolated environment for testing what AI agents actually do when you give them autonomy and tools.

And the Incubator for AI, which now sits in DCET. Who here is familiar with the Incubator? A few people. So the Incubator is essentially a spin-out of our program. It's a team that exists in the Department for Science and Technology that does what it says in the tent.

It incubates new AI solutions for usage across the public sector. Its original founding team, most of the technical team, were our fellows. And what's really cool now is not just seeing the work that they produced while they're there, but also the fact that we're able to collaborate with them when it comes to scaling up some of that work.

Here's one recent example. So this tool is Xtract. So a bunch of our people have worked on Xtract. It's a collaboration with DeepMind. It's built on Gemini, and it essentially digitizes large swathes of the planning application process, especially those bits that are currently largely handwritten, including handwritten, hand-drawn maps there as well.

This was unveiled by the Prime Minister at London Tech Week last year, and we're currently in the process of rolling it out to every local authority in England. As I said, only one in five planning applications are currently decided on time.

That has a massive impact on economic growth, and economic growth is basically the biggest challenge this country facesright now. So anything we can do to make a dent on that is really significant. I think aspirationally as well, this will hopefully get us to a place where more and more planning applications can be decided by AI automatically.

And then another interesting one, this is very current. The education gap is a big problem, not just in the UK, but elsewhere. Many of you will have read the papers about AI tutors. It's a really exciting moment, the prospect of being able to level the playing field somewhat and put world-class tutors in front of every child, regardless of their socioeconomic background.

But it's something that has to be done really carefully. So currently, we're working

on producing safeguards and evaluating various frontier models against benchmarks, not just to make sure that children can interact safely with these in a classroom environment, but also measuring them against various metrics. I think in this one, the relevant benchmark is the cognitive load placed on the student.

And then last but not least, the new kids on the block, Justice AI. Some of you might have been here for the Justice AI talk yesterday. Was anybody here?

Justice AI18:12

Guest18:21

Yes.

Eoin Mulgrew18:22

Okay, good. For most of you, this is new. Justice AI are a new team that have been set up in the MOJ, some of whom are over there. Hello.

Again, I wouldn't say a spin-off from the fellowship, that gives us way too much credit, but the founder of Justice AI is one of our former fellows, Dan James, who's doing brilliant work in there. And they are deploying forward-deployed engineers into prisons and into other parts of the criminal justice system.

So kind of taking an approach that we're doing in Number 10 with policy people and comms people and lawyers, but instead they're embedding with parole officers and prison wardens. And they're doing loads of really interesting work. I can't go into too much detail, but most of it is around using AI to stop the flow of drugs into prisons, to find efficiencies, where currently there's quite manual processes involving lots of people, and generally improving the security and safety within the prison system.

And one of those FDEs is over there. It's Will. Will's one of our current fellows. Sorry, Will, I've embarrassed you. I just added in your photo last night because I thought this was sort of a good point to end on.

Will is... So to give you an idea, a few months ago, Will was in California getting a tan. That's him outside HMP Wormwood on a rainy day. But yeah, Will dropped out of Harvard, started a company, got it into Y Combinator, made a bit of money, but wanted to come and work for us.

And that's his second week on the job, and he's standing outside a prison with the keys to that actual prison about to go in. And that is exactly what we're trying to do through this program. You've maybe done good stuff in industry.

Join Us20:11

Eoin Mulgrew20:11

That's brilliant. Come join us, and we'll give you the keys to the state and see what you can do. So yeah, look, it's really early days. It's sort of an experiment, what we're doing. But I think the proof points so far have been that actually small, elite teams can actually achieve quite a lot.

We're already saving money. We're already shipping new public services at an unprecedented speed. We're already reforming frontline public services, and we're already putting new AI capabilities into the hands of other teams at the top of government. So yeah, and surprise, surprise, this was a recruitment pitch.

We are hiring. So please do scan the QR code, and I'll be here the rest of the day if you want to come up and chat. Thank you.

I think I've got time for a couple of questions, possibly. Somebody can tell me if not. Cause.

Q&A21:10

Eoin Mulgrew21:20

Oh.

Guest21:26

So one of the earlier examples you showed, there was this chat to explain policies and not explain, but try different projections and see how they would behave. Do you have to deal with syncophancy with the fact that if you have a user that's not necessarily very well-versed in AI, just wants to hear what he wants to hear, can direct the tool towards, "Oh, look, I'm an absolutely brilliant mastermind.

My policy is going to be fantastic," despite the policy being actually bad, but the AI is basically.

Eoin Mulgrew22:01

Yeah. Should I cut income tax to 0%? You're absolutelyright. Yeah, that's a very real risk. So it's not something that we have encountered too much, but it's only because we have sort of red-teamed the models for that before we've put it into the hands of users.

We also provide quite a bit of upskilling. So a lot of the teams that we work with, we're creating tools for them. They're possibly lawyers. They're possibly sociologists, professors, whatever. So we do coach them on some of the risks that this presents.

But yeah, it's a good question.

Guest22:37

Thank you.

Jack22:41

Hello. Thank you very much for the speech. I'm Jack from Accenture. I think you beat our pitch for hiring, but we'll try as well. But my question on the policy side is, as you progress with this FDE type of model actually making an impact, how did you see, as you said, it's an early day experiment, how did you see you start to scale and essentially start dealing with the central government or with local governments, the different party lines, and so forth?

So it's kind of a real kind of governmental kind of human kind of stance. How do you see that's going to play out?

Eoin Mulgrew23:20

Yeah, 100%. So in terms of how this scales, so I suppose at the end I said, "Oh, you know, we can do quite a bit, and some of the people that join us then set up their new teams, that's great.

Is it enough to turn the oil tanker itself?" No. Possibly over time, but it would take a long time, and we need to solve these problems quicker than that. So we have been thinking about that. I think realistically, some of this stuff requires strategic intervention.

So we need to change the way the rest of government operates. Part of the reason, part of the bargain that we basically made with ministers was, "Let us take the shackles off. Let us set up a small team at the center that abides by different rules and use it as a proof point."

So this is almost like a pilot. I would like to see a lot of what we're doing become the norm, become BAU. This at the moment is basically a hack to get around the system. So we need to change that first of all.

I think also another, if we're talking really about scale, we talked about some fairly targeted use cases there. I think what we want to do over the next sort of 12 to 24 months is do more horizontal work, looking at processes.

So I should explain this to people as well. When you think of the civil service, you probably think of policy people working in some of the buildings around here that you can see out the window. That is a very small sliver of the civil service.

It's about 400,000 people. Most of them are call center operators. They're prison wardens. They're nurses, et cetera, et cetera. There are a lot of processes out there, whether that's like transcription, that every police person will tell you is the bane of their existence, or it will be those massive call centers in DWP, HMRC.

So yeah, if we want to dial up the ambition, I would like to see us going after more of those horizontal use cases that can be applied en masse across the system. Yeah, sorry, that was a bit of a long answer.

Guest25:24

I know it. I let you go off there, but that's good.

Eoin Mulgrew25:31

I think this is the last one. Yeah.

Guest25:34

Should I wait for the microphone?

Eoin Mulgrew25:39

You can shout if you want. It's up to you.

Guest25:45

Yeah, so I work for an ed tech company that, among other things, is making AI tutors. So I'd love to talk to you more about that. But one thing we find, I mean, the big problem with most kids is that you can make the best AI tutor in the world, but the real problem is motivation.

If you sit a kid down, you sit a 12-year-old down in front of a computer, they're going to do everything they can to avoid learning. So my question is, I mean, first of all, I'm just interested to know more about what the vision is of the government for this.

Is this going to be going into schools? Are kids going to be sitting down in front of computers and using it? And secondly, how do you solve that motivation problem?

Eoin Mulgrew26:20

Yeah, 100%. So at the moment, I think our plan is largely not to necessarily develop products that compete with yours, but rather set.

Guest26:30

Good news.

Eoin Mulgrew26:32

But rather set benchmarks and guardrails for how schools can then adopt whatever products they want. In terms of student uptake, though, that's not something we've done too much on yet. The test that you just saw was our initial testing, which has been done, I think, with 70 teachers who were then role-playing their pupils.

So actually, your experience might be quite valuable. So we should chat after.

Guest26:59

Just one more. I'm from Norway. So it's very great to see your ambitions, and I'm sure that many countries across Europe are doing exactly the same thing. Are you doing any form of collaboration with other countries, sharing ideas, et cetera?

Eoin Mulgrew27:16

Yeah, a bit. Norway, no, but if you've got contacts, I'd be very happy to chat to them. Yeah, we do a bit. There are a couple of teams that are sort of similar to what we're doing, albeit a little bit differently.

There are a couple of different initiatives underway in the US government, which are not a million miles away, stuff like Tech Force, parts of the US Digital Service, Singapore as well. We talk quite a bit with Singapore. But yeah, we could do more.

So yeah, if you have any contacts in the Norwegian government, I'd be very open to it.

Done.