Noor II
Being builtIn reviewSome careTier-1 Support Agent · Junior · IT & Digital Services
Sorts incoming service desk tickets, fixes the known ones and hands the rest to a person.
What you can do about it
Timesheet
Tasks and what they cost, per day, last 60 days
Performance review · 2026 Q2
Reviewed by Yasmin Farouk
Quality is below the bar for how much care this work needs. The checks it fails are the multi-step requests.
Could another model do this job better?
| Model | Quality | Cost per 1,000 tasks | Answers in | |
|---|---|---|---|---|
| Claude Opus 5 | 71 | $86.34 | 5.9s | Runs on this now |
| Claude Sonnet 5 | 64 | $51.80 | 4.3s | Cheaper and better |
| Claude Haiku 4.5 | 55 | $17.27 | 2.6s | |
| Falcon-70B (self-hosted) | 51 | $4.32 | 4.3s |
What it did
FailedThe system it depends on did not answer after three tries$0.46 · 1.0s · 22h ago
- 1. Read the request and worked out what kind of task it was
- 2. Called the system it depends on
- 3. Tried again three times, all three failed
- 4. Reported the failure back to whoever asked
FailedFound no documents it is allowed to read$0.34 · 2.3s · 3d ago
- 1. Read the request and worked out what kind of task it was
- 2. Called the system it depends on
- 3. Tried again three times, all three failed
- 4. Reported the failure back to whoever asked
DoneRan an approved data query and explained the result$0.19 · 2.5s · 3d ago
- 1. Read the request and worked out what kind of task it was
- 2. Found 6 possible passages, kept the 3 close enough to use
- 3. Drafted the answer and attached where each part came from
- 4. Checked its own claims back against those passages
DoneAnswered a policy question and cited the source clause$0.44 · 2.5s · 3d ago
- 1. Read the request and worked out what kind of task it was
- 2. Found 6 possible passages, kept the 3 close enough to use
- 3. Drafted the answer and attached where each part came from
- 4. Checked its own claims back against those passages
Stopped by a ruleThe model it runs on is not on the approved list$0.34 · 3.6s · 4d ago
- 1. Read the request and worked out what kind of task it was
- 2. Found a document to work from
- 3. A rule stopped it before it read the content
- 4. Stopped the task and recorded it
Handed backRequest outside what it was built for$0.15 · 4.9s · 4d ago
- 1. Read the request and worked out what kind of task it was
- 2. Searched everything it may read and found no good answer
- 3. Put together what it tried, what it found and what it needs
- 4. Passed it to the person who owns the work
FailedCould not sign in to the system it needed, its access had expired$0.53 · 1.5s · 4d ago
- 1. Read the request and worked out what kind of task it was
- 2. Called the system it depends on
- 3. Tried again three times, all three failed
- 4. Reported the failure back to whoever asked
DoneReviewed a contract against the clause playbook$0.25 · 3.4s · 5d ago
- 1. Read the request and worked out what kind of task it was
- 2. Found 6 possible passages, kept the 3 close enough to use
- 3. Drafted the answer and attached where each part came from
- 4. Checked its own claims back against those passages
DoneSummarised a 34-page document into a one-page brief$0.20 · 1.8s · 6d ago
- 1. Read the request and worked out what kind of task it was
- 2. Found 6 possible passages, kept the 3 close enough to use
- 3. Drafted the answer and attached where each part came from
- 4. Checked its own claims back against those passages
Felix II already does this job
Felix II holds "Tier-1 Support Agent" in IT & Digital Services and is doing that work every day. Job title, purpose, what it may do, what it may read, department and channels match. Extend that agent, or change this description until the two jobs are genuinely different.
Tier-1 Support Agent
IT & Digital Services · Yasmin Farouk · $5,697 a month across 2,318 tasks
Same job: job title, purpose, what it may do all match. Neither title narrows the job to a region or a business line, so both serve the same people.
What matched
- Job title
- Both hold "Tier-1 Support Agent"
- Purpose
- The job description is word for word the same
- What it may do
- Both can knowledge search, create ticket, resolve ticket
- What it may read
- Both read Service Desk Tickets, Operations Runbooks
- Channels
- Both reachable on Web portal, Teams, Email
- Department
- Both sit in IT & Digital Services
Identical permission envelope: same department, same permitted actions, same knowledge. That is a fact the platform granted rather than a judgement about the wording, so it counts even when the two job descriptions read nothing alike.
Tier-1 Support Agent: Group
IT & Digital Services · Yasmin Farouk · $10.4k a month across 3,602 tasks
Same job: job title, purpose, what it may do all match. Neither title narrows the job to a region or a business line, so both serve the same people.
What matched
- Job title
- Both hold "Tier-1 Support Agent"
- Purpose
- The job description is word for word the same
- What it may do
- Both can knowledge search, create ticket, resolve ticket
- What it may read
- Both read Service Desk Tickets, Operations Runbooks
- Channels
- Both reachable on Web portal, Teams, Email
- Department
- Both sit in IT & Digital Services
Identical permission envelope: same department, same permitted actions, same knowledge. That is a fact the platform granted rather than a judgement about the wording, so it counts even when the two job descriptions read nothing alike.
Measured against the agents doing the same work
3 agents in IT & Digital Services hold the same job. All 3 pairings are outright copies, not splits by region or business line. Built in Azure AI Foundry and Codex, so there was nowhere either builder could have looked.
| Measure | Felix II · keep2,318 tasks | Dimitri II3,602 tasks | Noor II956 tasks |
|---|---|---|---|
Cost per taskModel usage, running costs and your time, over tasks completed | $2.46● | $2.90 | $4.18 |
Tasks finished cleanlyRange allows for how few tasks some agents have run | 93.7% | 97.6%● | 84.5% |
Handed back to a personToo close to call: on 3,602 and 2,318 the two ranges still overlap | 8.1% | 7.3% | 16.7% |
How people rate itHeld toward the average across your agents of 4.2 until 8 people have rated it | 4.75 of 5● | 3.42 of 5 | 4.02 of 5 |
Checks passingToo close to call: on 198 and 159 the two ranges still overlap | 86% | 92% | 88% |
Review scoreQuality, accuracy, sticking to the source and safety from the last review, averaged | 94 | 97● | 68 |
Your time per 100 tasksApprovals, handbacks and review time, in hours of your people | 2.7 h● | 3.2 h | 4.6 h |
Value returned per dollarTime it saved, valued at the hourly rate of the people it saved it for, over full cost | 3.9×● | 0.6× | 1.7× |
OverallRanked on the cautious end of every range, so an agent has to do the work to win. Comparable inside this group only. | 92.5Strong | 85.9Strong | 63.2Moderate |
When it went wrong
Someone flagged an answer quoting a policy clause that does not exist. The record shows it found no source close enough to use and answered anyway.