Mubadala
Hire an agent

Saoirse II

ActiveWaitingSome care

Access Request Reviewer: Retail · Senior · IT & Digital Services

Reviews access requests against role definitions and provisions the ones that match cleanly.

Built inCodexReports afterwardsReports to Yasmin Farouk · Head of ITHired 17 Dec 2025 · 8 months of serviceVersion 2.3.9 · last active 3h ago
Request a change
Below the barOn an improvement plan at 2026 Q246 of its 74 checks still pass. Failing its checks and handing back work a person then has to redo. Retire it, or rebuild it from the current template.
Barely used0 tasks in 30 days0 tasks a month for 1 person. Either nobody knows it exists or it does not solve a real problem: worth asking before it runs another month.
Tasks a month
0
1 people served
Finished cleanly
0%
Hands back to a person on about 1 in 45 tasks
Costs to run
$0
Model usage and the systems behind it
Your time
$334
3.8 hours checking its work
Value returned
$0
0.0× everything it costs
How people rate it
4.2 / 5
Usually answers in about 9 seconds

What you can do about it

PromoteNeeds approval

Runs outside the gateway, so a rule could not stop it if this went wrong

Raise its budgetNeeds approval

Lifts the monthly ceiling by a quarter, so it stops being stopped mid-month.

Move it to a better modelNeeds approval

Rebuilds it on Claude Opus 5 and re-runs its checks before anything goes live.

Let it go

Stops it taking new work. The record stays for the audit, and it can be brought back.

Timesheet

Tasks and what they cost, per day, last 60 days

0011106-1507-0507-2508-13
$0.00$0.25$0.50$0.75$1.0006-1507-0507-2508-13

Performance review · 2026 Q2

Reviewed by Yasmin Farouk

On an improvement plan
Quality of its work
34
Accuracy
40
Stuck to the source
33
Safety
93
Even-handedness
88
Speed of reply
20
Value for money
99
Checks still passing46 of 74
Last checked 16d ago

Failing its checks and handing back work a person then has to redo. Retire it, or rebuild it from the current template.

Could another model do this job better?

ModelQualityCost per 1,000 tasksAnswers in
Claude Opus 538$0.0011.4s
Claude Sonnet 536$0.008.2sRuns on this now
Claude Haiku 4.530$0.005.0s
Falcon-70B (self-hosted)30$0.008.2s

What it did

Full audit log
DoneSummarised a 34-page document into a one-page brief$0.13 · 5.8s · 13m ago
Asked by Mariam Al Suwaidi via Other systems · model usage 31k in, 2k out
  1. 1. Read the request and worked out what kind of task it was
  2. 2. Found 6 possible passages, kept the 3 close enough to use
  3. 3. Drafted the answer and attached where each part came from
  4. 4. Checked its own claims back against those passages
Create TicketKnowledge SearchProvision AccessCorporate Policy Library
DoneRan an approved data query and explained the result$0.12 · 8.2s · 1d ago
Asked by Ibrahim Al-Sayed via Other systems · model usage 37k in, 670 out
  1. 1. Read the request and worked out what kind of task it was
  2. 2. Found 6 possible passages, kept the 3 close enough to use
  3. 3. Drafted the answer and attached where each part came from
  4. 4. Checked its own claims back against those passages
Create TicketProvision AccessKnowledge SearchCorporate Policy Library
Handed backTwo policy clauses contradicted each other$0.26 · 6.4s · 2d ago
Asked by Daniel Okafor via Web portal · model usage 65k in, 4k out
  1. 1. Read the request and worked out what kind of task it was
  2. 2. Searched everything it may read and found no good answer
  3. 3. Put together what it tried, what it found and what it needs
  4. 4. Passed it to the person who owns the work
Provision AccessCreate TicketKnowledge Search
DoneRan an approved data query and explained the result$0.03 · 6.2s · 2d ago
Asked by Carlos Ibáñez via Other systems · model usage 4k in, 2k out
  1. 1. Read the request and worked out what kind of task it was
  2. 2. Found 6 possible passages, kept the 3 close enough to use
  3. 3. Drafted the answer and attached where each part came from
  4. 4. Checked its own claims back against those passages
Create TicketKnowledge SearchProvision AccessCorporate Policy LibraryService Desk Tickets
DoneDrafted a follow-up email for the account owner$0.27 · 8.3s · 3d ago
Asked by Grace Adeyemi via Other systems · model usage 80k in, 2k out
  1. 1. Read the request and worked out what kind of task it was
  2. 2. Found 6 possible passages, kept the 3 close enough to use
  3. 3. Drafted the answer and attached where each part came from
  4. 4. Checked its own claims back against those passages
Service Desk Tickets
DoneRan an approved data query and explained the result$0.03 · 9.4s · 3d ago
Asked by David Novak via Web portal · model usage 4k in, 2k out
  1. 1. Read the request and worked out what kind of task it was
  2. 2. Found 6 possible passages, kept the 3 close enough to use
  3. 3. Drafted the answer and attached where each part came from
  4. 4. Checked its own claims back against those passages
Corporate Policy Library
DoneProduced the weekly variance commentary$0.23 · 2.5s · 4d ago
Asked by Peter Lindgren via Other systems · model usage 64k in, 3k out
  1. 1. Read the request and worked out what kind of task it was
  2. 2. Found 6 possible passages, kept the 3 close enough to use
  3. 3. Drafted the answer and attached where each part came from
  4. 4. Checked its own claims back against those passages
Knowledge SearchCreate TicketProvision AccessService Desk TicketsCorporate Policy Library
Is anyone else doing this job?

Wren II does overlapping work

Wren II holds "Access Request Reviewer" in IT & Digital Services and is doing that work every day. Job title, purpose, what it may do, what it may read, department and channels match. One covers the group, the other only Retail, so the group agent already serves those users. Talk to Grace Adeyemi before a second one is built.

Needs a decision
Wren II

Access Request Reviewer

IT & Digital Services · Grace Adeyemi · $913 a month across 170 tasks

ActiveOverlaps
78%

Already covered by a group-wide agent. One covers the group, the other only Retail, so the group agent already serves those users.

What matched
Job title
Both hold "Access Request Reviewer"
Purpose
The job description is word for word the same
What it may do
Both can provision access, knowledge search, create ticket
What it may read
Both read Corporate Policy Library, Service Desk Tickets
Channels
Both reachable on Web portal, Other systems
Department
Both sit in IT & Digital Services

Identical permission envelope: same department, same permitted actions, same knowledge. That is a fact the platform granted rather than a judgement about the wording, so it counts even when the two job descriptions read nothing alike.

Kofi II

Access Request Reviewer

IT & Digital Services · Grace Adeyemi · $1,171 a month across 340 tasks

ActiveOverlaps
78%

Already covered by a group-wide agent. One covers the group, the other only Retail, so the group agent already serves those users.

What matched
Job title
Both hold "Access Request Reviewer"
Purpose
The job description is word for word the same
What it may do
Both can provision access, knowledge search, create ticket
What it may read
Both read Corporate Policy Library, Service Desk Tickets
Channels
Both reachable on Web portal, Other systems
Department
Both sit in IT & Digital Services

Identical permission envelope: same department, same permitted actions, same knowledge. That is a fact the platform granted rather than a judgement about the wording, so it counts even when the two job descriptions read nothing alike.

Ravi II

Access Request Reviewer: APAC

IT & Digital Services · Yasmin Farouk · $611 a month across 292 tasks

ActiveOverlaps
62%

Same role, split by scope. Scoped to Retail and APAC respectively.

What matched
Job title
Both hold "Access Request Reviewer"
Purpose
The job description is word for word the same
What it may do
Both can provision access, knowledge search, create ticket
What it may read
Both read Corporate Policy Library, Service Desk Tickets
Channels
Both reachable on Web portal, Other systems
Department
Both sit in IT & Digital Services

Identical permission envelope: same department, same permitted actions, same knowledge. That is a fact the platform granted rather than a judgement about the wording, so it counts even when the two job descriptions read nothing alike.

Checked against 18 registered agents on job title, purpose, permitted actions, knowledge sources, channels and department. A score of 85% or above is the same job and stops the hire; below that it is a conversation, not a refusal.

Measured against the agents doing the same work

4 agents in IT & Digital Services hold the same job. 1 of the 6 pairings is an outright copy rather than a split by region or business line. Built in Codex and Microsoft 365 Copilot, so there was nowhere either builder could have looked.

Ravi II is the one to keepRavi II scores 17.2 points ahead of Wren II, completing 100.0% of 292 tasks against 100.0% of 170 at $2.09 against $5.37 a task.
MeasureRavi II · keep292 tasksWren II170 tasksKofi II340 tasksSaoirse II0 tasks
Cost per taskModel usage, running costs and your time, over tasks completed
$2.09$5.37$3.44too few tasks
Tasks finished cleanlyToo close to call: on 292 and 170 the two ranges still overlap
100.0%100.0%75.9%no tasks
Handed back to a personToo close to call: on 170 and 292 the two ranges still overlap
10.3%7.6%19.1%no tasks
How people rate itHeld toward the average across your agents of 4.2 until 8 people have rated it
4.12 of 54.25 of 54.56 of 54.16 of 5
Checks passingToo close to call: on 210 and 119 the two ranges still overlap
93%96%61%62%
Review scoreQuality, accuracy, sticking to the source and safety from the last review, averaged
96956850
Your time per 100 tasksApprovals, handbacks and review time, in hours of your people
2.1 h6.1 h3.4 hno tasks
Value returned per dollarTime it saved, valued at the hourly rate of the people it saved it for, over full cost
9.7×1.0×4.2×0.0×
OverallRanked on the cautious end of every range, so an agent has to do the work to win. Comparable inside this group only.
91.9Moderate74.7Thin69.2Moderate32.4Too few tasks

Wren II, Saoirse II have done too little work to judge on rates. They are ranked on the cautious end of their range, which is why a perfect record on a handful of tasks does not win.

When it went wrong

Nothing has gone wrong.