How AI wants productivity apps
Productivity Apps is decided on affinity: Google Calendar indexes 115 there against ClickUp at 83, a 32-point spread. Google Calendar holds the strongest overall profile at 107 — quietly trusted. But the brand assistants name most often is Apple Notes, not the category's market leader: being desired and being named are not the same measurement. Across these 12 brands, affinity tracks being named most closely (rank correlation 0.74).
12 brands · 60 AI answers · 156 dimension scores · measured, never estimated
Productivity Appsupset
| # | Brand | Attraction | Affinity | Attachment | Desire shape | Named in |
|---|---|---|---|---|---|---|
| 1 | Google Calendar | 101 | 115 | 113 | Quietly trusted | 23% |
| 2 | Todoist | 106 | 108 | 108 | Complete desire | 23% |
| 3 | Notion | 114 | 102 | 95 | Magnetic but unbought | 37% |
| 4 | Apple Notes | 97 | 109 | 108 | Quietly trusted | 58% |
| 5 | Obsidian | 112 | 96 | 94 | Magnetic but unbought | 23% |
| 6 | TickTick | 102 | 96 | 104 | Habit, not love | 17% |
| 7 | Trello | 101 | 96 | 94 | Magnetic but unbought | 7% |
| 8 | ClickUp | 105 | 83 | 94 | Magnetic but unbought | 8% |
| 9 | Google Keep | 86 | 103 | 108 | Habit, not love | 53% |
| 10 | Microsoft OneNote | 90 | 103 | 102 | Quietly trusted | 27% |
| 11 | Asana | 98 | 92 | 88 | Magnetic but unbought | 8% |
| 12 | Evernote | 86 | 96 | 92 | Fading | 17% |
Indexed against the productivity apps average of 12 brands (100 = the category), measured across 60 AI answers in the latest weekly run.
Driver by driver
Where the category actually separates. A wide spread means assistants describe these brands very differently on that driver; a narrow one means it is table stakes.
Affinity
The widest field of the three: this is where the category separates.
Attraction
Real separation, but not the category’s defining contest.
Attachment
The tightest of the three: most brands land close together here.
The shapes AI leaves these brands in
Three indices make a profile, and profiles fall into recognisable shapes. A brand that is magnetic but unbought has a different problem from a habit nobody loves — and a different thing to build next.
What assistants actually talk about
The sixteen dimensions underneath the drivers, scored 0–10 on how the answers describe each brand. Spread is the distance between the strongest and weakest brand: the widest are where a brand can still separate itself, the narrowest are the price of entry.
Where brands separate
| Dimension | Driver | Category avg | Strongest | Spread |
|---|---|---|---|---|
| Value & Pricing | Attachment | 7.8 | Apple Notes 9.3 | 3.3 |
| Boldness & Creativity | Attraction | 6.8 | Notion 8.5 | 3.0 |
| Ease of Use & Onboarding | Attachment | 7.8 | Apple Notes 9.0 | 3.0 |
| Heritage & Authority | Affinity | 7.7 | Google Calendar 9.0 | 3.0 |
| Differentiation vs Sameness | Attraction | 7.4 | Notion 8.5 | 2.5 |
Where everyone reads alike
| Dimension | Driver | Category avg | Strongest | Spread |
|---|---|---|---|---|
| Quality & Performance | Attraction | 8.2 | Google Calendar 9.0 | 1.5 |
| Trust & Credibility | Affinity | 8.2 | Google Calendar 9.0 | 2.0 |
| Reliability & Consistency | Attachment | 8.2 | Google Calendar 9.0 | 2.0 |
| Range & Availability | Attachment | 8.1 | ClickUp 9.0 | 2.0 |
| Lifestyle Fit & Identity | Affinity | 6.9 | Apple Notes 8.0 | 2.0 |
Does desire track being recommended?
Rank correlation between a brand's driver index and the share of answers naming it, across the 12 indexed brands in this category. A relationship inside one run's answers — descriptive, never causal, and easily moved by a single brand at this sample size.
Method & limitations
Every week, each category runs a fixed set of unbranded consumer prompts against ChatGPT, Claude and Gemini. A separate model reads the answers and scores every brand named in them across 16 perception dimensions, 0–10. Those dimensions are grouped into the three Drivers of Desire from the Havas Science of Desire framework, and each brand's driver score is indexed against the average of every scored brand in its category — 100 is the category, 115 is fifteen per cent above it.
This measures how AI assistants describe brands in their answers. It is not a survey of people, and it is not an endorsement: the scores are a reading of machine-generated text at the time of the run, and relationships between drivers and how often a brand is named are correlational. A brand needs at least six scored dimensions in a run to be indexed, and a category needs at least four such brands to get a board. See the full methodology.
Other categories
Where next
This edition measures what AI wants from these brands. These answer what comes next — which brands AI names, which websites feed those answers, and how any of it is counted.