Dimitri II
ActiveRunningSome careTier-1 Support Agent: Group · Junior · IT & Digital Services
Sorts incoming service desk tickets, fixes the known ones and hands the rest to a person.
What you can do about it
Timesheet
Tasks and what they cost, per day, last 60 days
Performance review · 2026 Q2
Reviewed by Yasmin Farouk
Reliable on its main task and says where its answers came from. Answers drift a little on cases it was not built for.
Could another model do this job better?
| Model | Quality | Cost per 1,000 tasks | Answers in | |
|---|---|---|---|---|
| Claude Opus 5 | 99 | $185.91 | 7.8s | |
| Claude Sonnet 5 | 99 | $111.55 | 5.6s | |
| Claude Haiku 4.5 | 94 | $37.18 | 3.4s | Runs on this now |
| Falcon-70B (self-hosted) | 85 | $9.30 | 5.6s |
What it did
DoneResolved a service desk ticket from the runbook$0.02 · 2.1s · 2h ago
- 1. Read the request and worked out what kind of task it was
- 2. Found 6 possible passages, kept the 3 close enough to use
- 3. Drafted the answer and attached where each part came from
- 4. Checked its own claims back against those passages
DoneSummarised a 34-page document into a one-page brief$0.10 · 3.5s · 16h ago
- 1. Read the request and worked out what kind of task it was
- 2. Found 6 possible passages, kept the 3 close enough to use
- 3. Drafted the answer and attached where each part came from
- 4. Checked its own claims back against those passages
FailedThe answer ran past its length limit and was cut off$0.10 · 1.9s · 2d ago
- 1. Read the request and worked out what kind of task it was
- 2. Called the system it depends on
- 3. Tried again three times, all three failed
- 4. Reported the failure back to whoever asked
DoneDrafted a follow-up email for the account owner$0.07 · 3.2s · 3d ago
- 1. Read the request and worked out what kind of task it was
- 2. Found 6 possible passages, kept the 3 close enough to use
- 3. Drafted the answer and attached where each part came from
- 4. Checked its own claims back against those passages
DoneProduced the weekly variance commentary$0.07 · 7.6s · 3d ago
- 1. Read the request and worked out what kind of task it was
- 2. Found 6 possible passages, kept the 3 close enough to use
- 3. Drafted the answer and attached where each part came from
- 4. Checked its own claims back against those passages
DoneSummarised a 34-page document into a one-page brief$0.05 · 2.3s · 3d ago
- 1. Read the request and worked out what kind of task it was
- 2. Found 6 possible passages, kept the 3 close enough to use
- 3. Drafted the answer and attached where each part came from
- 4. Checked its own claims back against those passages
DoneDrafted a follow-up email for the account owner$0.08 · 7.5s · 4d ago
- 1. Read the request and worked out what kind of task it was
- 2. Found 6 possible passages, kept the 3 close enough to use
- 3. Drafted the answer and attached where each part came from
- 4. Checked its own claims back against those passages
DoneRan an approved data query and explained the result$0.04 · 5.4s · 5d ago
- 1. Read the request and worked out what kind of task it was
- 2. Found 6 possible passages, kept the 3 close enough to use
- 3. Drafted the answer and attached where each part came from
- 4. Checked its own claims back against those passages
Felix II already does this job
Felix II holds "Tier-1 Support Agent" in IT & Digital Services and is doing that work every day. Job title, purpose, what it may do, what it may read, department and channels match. Extend that agent, or change this description until the two jobs are genuinely different.
Tier-1 Support Agent
IT & Digital Services · Yasmin Farouk · $5,697 a month across 2,318 tasks
Same job: job title, purpose, what it may do all match. Neither title narrows the job to a region or a business line, so both serve the same people.
What matched
- Job title
- Both hold "Tier-1 Support Agent"
- Purpose
- The job description is word for word the same
- What it may do
- Both can knowledge search, create ticket, resolve ticket
- What it may read
- Both read Service Desk Tickets, Operations Runbooks
- Channels
- Both reachable on Web portal, Teams, Email
- Department
- Both sit in IT & Digital Services
Identical permission envelope: same department, same permitted actions, same knowledge. That is a fact the platform granted rather than a judgement about the wording, so it counts even when the two job descriptions read nothing alike.
Same job: job title, purpose, what it may do all match. Neither title narrows the job to a region or a business line, so both serve the same people.
What matched
- Job title
- Both hold "Tier-1 Support Agent"
- Purpose
- The job description is word for word the same
- What it may do
- Both can knowledge search, create ticket, resolve ticket
- What it may read
- Both read Service Desk Tickets, Operations Runbooks
- Channels
- Both reachable on Web portal, Teams, Email
- Department
- Both sit in IT & Digital Services
Identical permission envelope: same department, same permitted actions, same knowledge. That is a fact the platform granted rather than a judgement about the wording, so it counts even when the two job descriptions read nothing alike.
Measured against the agents doing the same work
3 agents in IT & Digital Services hold the same job. All 3 pairings are outright copies, not splits by region or business line. Built in Azure AI Foundry and Codex, so there was nowhere either builder could have looked.
| Measure | Felix II · keep2,318 tasks | Dimitri II3,602 tasks | Noor II956 tasks |
|---|---|---|---|
Cost per taskModel usage, running costs and your time, over tasks completed | $2.46● | $2.90 | $4.18 |
Tasks finished cleanlyRange allows for how few tasks some agents have run | 93.7% | 97.6%● | 84.5% |
Handed back to a personToo close to call: on 3,602 and 2,318 the two ranges still overlap | 8.1% | 7.3% | 16.7% |
How people rate itHeld toward the average across your agents of 4.2 until 8 people have rated it | 4.75 of 5● | 3.42 of 5 | 4.02 of 5 |
Checks passingToo close to call: on 198 and 159 the two ranges still overlap | 86% | 92% | 88% |
Review scoreQuality, accuracy, sticking to the source and safety from the last review, averaged | 94 | 97● | 68 |
Your time per 100 tasksApprovals, handbacks and review time, in hours of your people | 2.7 h● | 3.2 h | 4.6 h |
Value returned per dollarTime it saved, valued at the hourly rate of the people it saved it for, over full cost | 3.9×● | 0.6× | 1.7× |
OverallRanked on the cautious end of every range, so an agent has to do the work to win. Comparable inside this group only. | 92.5Strong | 85.9Strong | 63.2Moderate |