Measuring AI adoption across your organization

Ready to Accelerate AI Adoption Across Your Team?

See How It Works

One Dashboard for Claude, Copilot, Gemini, and ChatGPT

One unified AI tool dashboard for Claude, Copilot, Gemini and ChatGPT. Worklytics MeasureAI tracks AI adoption, engagement, productivity, cost and impact.

TL;DR

  • Claude, ChatGPT, Copilot, and Gemini each give admins their own usage report. Those four reports do not add up to one adoption number, because each one counts something different. Claude counts requests and spend per person. ChatGPT counts messages. Copilot counts how many people used it inside each Office app. Gemini counts how often people used each AI feature and how close they are to their usage cap.
  • They also compare against different groups of people. Copilot divides active users by the people who have Copilot switched on. ChatGPT shows seats bought next to seats switched on. Claude shows active members against seats assigned. Gemini shows active users against everyone whose licence includes it. Averaging those four percentages does not give you a real adoption rate.
  • A consolidated AI dashboard has to line up four things before it can compare anything: who the person is, which group of people you measure them against, what time period you use, and what counts as one use.
  • Worklytics builds that layer on top of the connections each AI tool already offers, then goes past usage into how effectively AI is used and whether output actually changed. See the AI adoption dashboard for the combined view and Benchmark to compare against similar companies.
  • Counting logins tells you people showed up. Linking AI use to meeting load, focus time, and each manager’s team turns a tool report into a real productivity measurement.

What Is a Consolidated AI Dashboard?

A consolidated AI dashboard is one report that pulls usage, cost, and results from every AI tool a company runs, and puts those numbers in the same terms so you can compare them, add them up, and rank them. It sits on top of the reports each AI tool already gives you, rather than replacing them.

The difference that matters is between collecting and comparing. Downloading four spreadsheets into one folder is collecting. A consolidated AI dashboard only becomes useful once the same employee, the same time period, and the same definition of a use apply across all four tools. Everything below is about that second step, because it is where most in-house attempts get stuck.

Why Four Separate AI Reports Do Not Produce One Adoption Number

Anthropic, OpenAI, Microsoft, and Google all give admins a usage report, a filter for looking at specific teams, and a way to download the data. The problem is not that the reports are missing. The problem is that their numbers cannot be compared.

Each report was built around how that company charges you, and that decided what gets counted. Anthropic charges for how much you use, so Claude reports requests, tokens (the units of text an AI model processes), and dollars spent. OpenAI charges per person, so ChatGPT reports messages and how many people have started using it. Microsoft charges per person per app, so Copilot reports how many people used it inside Word, Outlook, Teams, and the rest. Google limits how much AI each subscription tier allows, so Gemini reports feature use and how close people are to their limit.

So when an executive asks what the company AI adoption rate is, they get four answers that cannot be averaged, added, or ranked. A department at 70% on Copilot and 20% on Claude is not 45% adopted. Those two numbers describe different sets of people, over different time periods, counting different things.

What Each AI Tool’s Own Report Actually Measures

The table below compares the four reports on the points that decide whether their numbers can be combined. Every value comes from the official documentation each company publishes for admins.

AI tool What it counts as one use Who it compares against How current the data is Where team labels come from
Claude Enterprise Requests, tokens used, conversations, pull requests created, and an estimate of time saved People active in the period, against the seats assigned to them Updates daily; spend figures run one day behind Groups set up inside Claude; Enterprise plans can also pull the data automatically
ChatGPT Enterprise Messages, split into ordinary chats, custom GPTs, tools, and projects Active users, against both seats bought and seats switched on Every 1 to 24 hours, usually within 6 to 12 Groups synced from your staff directory; teams under 10 people are hidden in task reports
Microsoft 365 Copilot People who used Copilot inside each Microsoft 365 app; AI agents are reported separately Active users divided by the people who have Copilot switched on Within 48 hours, covering the last 28 days by default Microsoft Entra, the company staff directory, shown through Viva Insights
Google Gemini Times each AI feature was used in each Workspace app, plus high, medium, low, and zero use bands Active users, against everyone whose licence includes Gemini Changes to your team structure take up to 72 hours to appear Google Workspace organizational units and groups; new groups take up to 5 days to show

Three things in that table create most of the cleanup work. First, Copilot and ChatGPT report on a list of specific employees that IT already manages, while Claude reports on how much was spent, which finance manages. Second, the reports update at different speeds, so a Monday morning snapshot of all four covers four different stretches of time. Third, the team labels come from different staff directories, so Engineering in Microsoft, Engineering in ChatGPT, and Engineering in Google Workspace often contain different people. For step-by-step guides on each tool, see tracking Claude Enterprise usage, exporting ChatGPT Enterprise usage data, and tracking Copilot use.

The Four Things a Consolidated AI Dashboard Has to Line Up

Combining these reports is not just a matter of moving data around. Pulling four sets of data into one place gives you four disconnected tables. The work that makes a consolidated AI dashboard usable happens after the data arrives, and it runs in this order.

1. The person: match every account to one employee

One engineer can have a Claude seat under a work email, a GitHub Copilot licence under a GitHub username, a Gemini licence under a Google Workspace team, and a ChatGPT seat created automatically by your staff directory. Until those four records point to the same employee, you cannot work out what AI costs per person or who is actually using it. Worklytics matches accounts against your HR system, which also brings in department, tenure, level, and who each person reports to.

2. The comparison group: pick one set of people and stick to it

Each AI tool measures against whichever group it charges you for. To get a number you can compare, pick one group, usually everyone in a department, and measure every tool against it. This changes the answer a lot. A tool used by 80% of its 60 licence holders reaches 12% of a 400-person department. Only the second number is useful when deciding what to renew.

3. The time period: put everything on the same calendar

Copilot reports the last 28 days by default. Claude spend runs a day behind. Gemini takes up to 72 hours to reflect team changes. Adding everything up weekly, on a fixed day, smooths those differences out. Reporting faster than that means the slowest tool quietly sets how current your whole dashboard really is.

4. The unit: turn different activities into one shared measure

Messages, requests, and feature uses are not the same thing, so you cannot add them together. Two measures do work across all four tools: how many days a week someone uses AI, and how many times they use it on the days they do. Days per week shows whether it has become a habit. Uses per active day shows how deeply they rely on it. Together they separate someone who opened a tool once from someone who works inside it, without depending on how any one tool defines an event.

Sample report: one table covering every AI tool, measured against the same set of employees using the same two depth measures.

How each measure is calculated is published in the Worklytics Data Dictionary, so your analysts can check where a number came from instead of trusting the label. Teams who want the combined data inside their own reporting tools can send it through DataStream and join it to the HR and delivery data they already hold.

How to Measure How Deeply Teams Use Claude, Copilot, Gemini, and ChatGPT

Counting who logged in is a licensing number. It tells you someone signed in once. Depth tells you whether the tool changed how they work, and depth is what predicts whether the licence is worth renewing.

Plotting days per week against uses per active day sorts every tool and every team into four groups, each needing a different response:

  • Core: used often and used deeply. Protect the licence, and copy what these teams do into other teams.
  • Habitual: used often but only for quick questions. The tool is part of the day but is being used as a search box. Training on longer, more specific prompts helps here. More seats will not.
  • Specialist: used rarely but deeply. This is the right pattern for tools built for one job. Coding assistants usually sit here, and that is fine.
  • Exploring: used rarely and shallowly. Either it overlaps with another tool you already pay for, or nobody was ever shown how to use it. This is where cutting licences pays for itself.
Four-box chart plotting AI tools by days per week against uses per active day, sorting them into core, habitual, specialist, and exploring

Sample report: tools sorted by how often and how deeply they are used, which separates a wasted licence from one that is simply specialized.

The same chart applied to departments instead of tools shows where training effort pays off most. A department using AI on few days but going deep each time already knows how to use it and just needs the habit. A department using it most days but only briefly has both a habit and a technique gap, and that responds better to prompt examples from colleagues than to training from the AI company.

Sample report: how deeply each department uses AI, separating a technique gap from a habit gap.

Looking across all tools at once answers a question no single tool can. The average number of AI tools each active person uses shows whether employees are settling on one main assistant or spreading thin across several that overlap. A high share of people using only one tool means the rest of what you pay for is sitting unused.

Sample report: measures that only exist once all the tools sit in one place.

How a Consolidated AI Dashboard Shows What AI Is Being Used For

How much a tool is used does not tell you what it is worth. A thousand messages a week spent rewriting emails and a thousand spent generating code return very different value and justify very different budgets.

Worklytics sorts AI activity into types of work: coding, research, analysis, summarizing, drafting, and writing emails. It does this from records of what happened, not from reading what people typed. The same set of categories is applied to Claude, Copilot, Gemini, and ChatGPT at once. That gives you a breakdown of what each tool is used for, which shows overlap straight away. Two tools doing near-identical work in the same department are competing for the same job, and one of them is a candidate for cancellation. The tools Worklytics connects to are listed on the integrations page.

Sample report: what each tool is used for, which turns overlapping licences into a question you can actually decide.

Steadiness is the second signal worth watching. A tool used in bursts behaves differently from one used at a steady weekly rhythm, even when the totals match. Bursts rarely change how people work, so they rarely change results.

Sample report: how steadily each tool is used over 14 weeks, separating lasting habits from short-lived spikes.

How to Compare Your AI Adoption Against Similar Companies

An internal number is hard to read on its own. A 22% adoption rate is poor if similar companies sit at 45% and strong if they sit at 8%. ChatGPT includes a comparison against an industry average for a few of its own measures, which is useful as far as it goes and covers only that one tool.

Worklytics Benchmark compares overall AI adoption, weekly usage, how many different AI agents are in use, AI use in meetings, and adoption by department against similar companies, across every tool rather than one. Seeing where you rank matters more than seeing an average. A company sitting in the bottom half on overall adoption but the top fifth on sales-team AI use does not have a company-wide problem. It has some teams far ahead and others far behind, and those two situations need different budgets.

Illustrative example: where a company ranks against similar organizations on five measures of AI adoption.

How to Track AI Cost by Tool, by Team, and by Result

Most companies now pay for AI in three different ways at once. Anthropic charges for how much you use and gives you a daily breakdown of spend per person and per model. OpenAI charges per person. Microsoft mixes a per-person licence with credits you spend as you go. Google limits AI use by which subscription tier you bought.

To combine those, you convert each one to the same figure: cost per active user per month, charged back to the department that ran it up. That single step answers three questions no individual tool report can.

  1. Which departments spend more on AI than their headcount would suggest, and whether that is deliberate.
  2. Which tools have more paid licences than weekly users. That is your list of licences to cut.
  3. Which tools produce the most real work per dollar. That is your list of licences to protect.
Sample report: AI spend charged back to departments, which shows concentration that an invoice hides.

Engineering usually costs the most, because coding assistants run constantly and consume a lot. That on its own is not a problem. The case worth acting on is a tool with many paid licences and few weekly users, where you can recover the difference at renewal without affecting anyone who is actually using it.

Sample report: cost against estimated value by tool. The value figures come from a time-saved calculation, and the inputs are shown next to the result.

Estimates like these rest on assumptions, and those assumptions should be visible rather than hidden. Worklytics shows the inputs behind its time-saved figures so your finance team can swap in its own salary costs. The AI ROI calculator runs the same sums against your own headcount and licence costs before you commit to anything.

How to Measure AI’s Effect on Productivity Without Relying on Surveys

ChatGPT measures its effect on a company through optional surveys inside the product, and OpenAI states plainly that these are rough indicators and do not prove AI caused any gain. Claude reports an estimated time saved, worked out from activity it can see. Both are reasonable within their own limits, and neither can see work that happens outside their own product.

Measuring behavior covers the rest. Worklytics tracks three stages that turn a usage report into a real measurement: adoption (who is using AI), proficiency (how much of the work is AI-assisted), and leverage (whether people are getting more done in a day). The full method is on the AI adoption dashboard page.

The Worklytics measurement model: adoption, proficiency, and leverage, each answering a different question.

Time-saved figures are only believable when tied to specific kinds of work rather than claimed for the whole company. Code generation and data analysis consistently save the most per person, because both replace long stretches of work rather than single steps. Writing emails saves the least, because the task was short to begin with.

Sample report: hours saved per active user per week, broken down by task instead of claimed company-wide.

The other half of the picture is what has not happened yet. Measuring which repetitive work is still being done without AI puts the remaining opportunity in the same units as the value you have already captured. That is what lets you compare a case for more training against a case for more licences.

Sample report: the opportunity not yet taken, measured the same way as the value already captured.

How a Consolidated AI Dashboard Stays Privacy-Safe

Bringing four tools together makes the resulting data more sensitive than any one of them on its own, so the rules around it have to be stricter than any single tool requires.

  • Records only, never content. How often, when, which tool, and what type of work are analyzed. What people typed, what the AI replied, and what was in their files are not collected at all.
  • Names replaced with codes before analysis. The swap happens as the data is collected, so the reporting system never holds anything that identifies a person directly.
  • Minimum team sizes. Small teams are hidden rather than shown, so a department view can never come down to one identifiable person. ChatGPT does the same thing by hiding teams under 10 people in its task reports.
  • Team-level reporting as standard. Departments and teams, not individuals, are what adoption and productivity reports are built on.

How this is built, and the layer that enforces it, is documented on the Worklytics privacy page. This matters commercially as well as ethically. Employee representative groups in several European countries have grounds to challenge AI monitoring where a company cannot show that content is left behind at the point of collection, rather than filtered out later in the report.

A 30-Day Plan for Setting Up a Consolidated AI Dashboard

  1. Days 1 to 5. Connect the AI tools you already pay for, and connect your HR system. The HR link is what lets you match accounts to people and compare against headcount, so it comes first rather than later.
  2. Days 6 to 10. Settle the comparison group. Agree one definition of adoption, measured against department headcount, and stop putting each tool’s own adoption figure in front of executives.
  3. Days 11 to 20. Sort by depth. Run the four-box chart across tools and departments, and split the list of licences to cut from the list of teams to train. These two lists point in opposite budget directions and are often confused.
  4. Days 21 to 30. Connect it to results. Link AI use to meeting load, focus time, and delivery data to set a starting point. You cannot measure a change without a before, so this is worth starting before anyone asks for the ROI number.

Connecting the tools usually produces first numbers within a week. Reading a trend takes about 30 days, because week-to-week swings in AI use are big enough to mislead before then.

Sample report: the company-level summary you get once person, group, time period, and unit all line up.

Frequently Asked Questions

What is a consolidated AI dashboard?

One report that combines usage, cost, and results from every AI tool a company runs, and puts those numbers in the same terms so they can be compared and added up. It is different from the report inside each AI tool, which covers only that product in its own units, and different from a software spend tracker, which standardizes cost but not how deeply a tool is used.

Can I build a consolidated AI dashboard myself?

The data is available from all four tools. Claude offers an automatic data feed on Enterprise plans. ChatGPT lets admins download spreadsheets on request, with a separate feed for detailed records and another one covering only its coding product. Microsoft makes Copilot usage available through its developer tools. Gemini exports from the Google Admin console. The hard part is not getting the data out. It is matching accounts across four staff directories, agreeing one comparison group, lining up the time periods, and keeping it all working as each company changes what it reports.

How often do these reports change?

Often enough that a home-built version needs someone looking after it. In the first half of 2026 alone, Google added new Gemini usage and limit reports to the Admin console, Microsoft cut its Copilot refresh time to 48 hours, moved its default view from 30 days to 28, and split AI agent figures into a separate report, and Anthropic added cost breakdowns by team and by person to Claude. Anything you build in-house has to be updated each time.

What is the difference between a consolidated AI dashboard and a software spend tracker?

A spend tracker standardizes cost and licence data. It tells you what you pay and how many licences sit unused. A consolidated AI dashboard also tells you how deeply the tools are used, for what kind of work, and whether output changed. The two overlap on unused licences and part ways everywhere else. A renewal decision usually needs both.

Does a consolidated AI dashboard read employee prompts?

A privacy-first one does not. Worklytics analyzes records of activity: how often, when, which tool, and what type of work it looks like based on surrounding signals. What people typed and what the AI replied are never collected. Any provider that needs to read prompts in order to categorize usage carries a very different set of privacy obligations, and that is worth confirming before you buy.

How many AI tools does a company usually need to combine?

More than the finance list suggests. Once you count coding assistants, meeting notetakers, and the AI features built into Microsoft 365 and Google Workspace licences, the number of places AI is actually being used is normally higher than the number of AI contracts signed, because some arrived inside licences bought for other reasons. Work the number out from usage data, not from the contract list.

Which single measure best shows AI adoption is working?

Uses per active day, tracked over time and split by department. Weekly active users tells you people turned up. Uses per active day tells you the tool became part of the work. Rising user numbers with flat depth means the rollout reached people without changing how they work, which is the most common way AI programs quietly fail.

Can we export the data to our own systems?

Yes. Worklytics DataStream sends the combined AI usage data to your own data warehouse or reporting tool, so you can join it to the HR, finance, and delivery data you already hold.

Where This Leaves AI Measurement

Anthropic, OpenAI, Microsoft, and Google will keep improving their admin reports, and each will keep measuring its own product in its own units, because that is what their pricing requires. None of them is in a position to compare your whole set of AI tools against each other.

That has to sit above them, joined to the staff and work data you already own. Worklytics builds it from data your company already holds, covering AI adoption, engagement, productivity, and results, so the answer to how your AI investment is doing is one number you can explain, rather than four that cannot be added together.

Request a demo

Schedule a demo with our team to learn how Worklytics can help your organization.

Book a Demo