Microsoft Copilot is used like a workplace tool on desktops and a personal companion on phones: mobile conversations skew heavily to health at all hours, while desktop queries concentrate on work and technology during business hours, with clear daily and weekly topic cycles.
Citation observations
Cumulative provider counts captured on specific dates; providers are never combined.
No provider observation is available for this paper.
Missing data, not a zero citation count.
We analyze 37.5 million deidentified conversations with Microsoft's Copilot between January and September 2025. Unlike prior analyses of AI usage, we focus not just on what people do with AI, but on how and when they do it. We find that how people use AI depends fundamentally on context and device type. On mobile, health is the dominant topic, which is consistent across every hour and every month we observed - with users seeking not just information but also advice. On desktop, the pattern is strikingly different: work and technology dominate during business hours, with "Work and Career" overtaking "Technology" as the top topic precisely between 8 a.m. and 5 p.m. These differences extend to temporal rhythms: programming queries spike on weekdays while gaming rises on weekends, philosophical questions climb during late-night hours, and relationship conversations surge on Valentine's Day. These patterns suggest that users have rapidly integrated AI into the full texture of their lives, as a work aid at their desks and a companion on their phones.
Summary
Main Finding
How and when people use conversational AI depends strongly on device and time: Copilot functions as a workplace productivity colleague on desktops during business hours, and as a constant personal confidant on mobile—especially for health and wellbeing—across all hours. Temporal rhythms (hour of day, weekdays vs weekends, seasonal/holiday spikes) create distinct use modes with clear economic and design implications.
Key Points
- Dataset and scale: analysis of 37.5 million deidentified Copilot conversations sampled daily from Jan 7–Sep 25, 2025 (≈260 days; ≈144,000 conversations/day).
- Device split:
- Mobile: dominated by Health & Fitness (Health/Fitness + Searching is the top topic-intent pair across every hour and month); mobile usage is comparatively stable month-to-month.
- Desktop: dominated by Technology and Work & Career, with Work & Career overtaking Technology between 8 a.m. and 5 p.m.; desktop patterns are more volatile (more topic-intent churn).
- Three interaction “modes” identified:
- The Workday (≈8 a.m.–5 p.m.): desktop-centred productivity (Work & Career, Technology, Education, Science).
- The Constant Companion: mobile-centred, persistent personal topics (notably health/advice) irrespective of hour.
- The Introspective Night: late-night increases in Religion & Philosophy, Personal Growth, and relationships.
- Temporal rhythms:
- Programming spikes on weekdays; gaming spikes on weekends.
- Holiday/seasonal effects: e.g., Relationships rises around Valentine’s Day; social/cultural topics gained prevalence over the period (shift from programming toward society/culture by September).
- Intent mix: the most common intents are Searching, Advice, Creating, Learning, and Technical Support.
- Methodological notes: conversations are labeled by automated, privacy-preserving “eyes‑off” classifiers into ~30 topical categories and ~11 intents (≈300 topic-intent pairs); ambiguous or short chats receive “No Topic/No Intent”.
- Data exclusions and privacy: enterprise-authenticated traffic is excluded; data are deidentified and processed per Microsoft privacy policies.
Data & Methods
- Source: Random daily samples of Copilot user interactions (consumer accounts only), Jan 7–Sep 25, 2025.
- Sample size: ~37.5M conversations total; sampled ~144k conversations per day.
- Privacy: automatic PII scrubbing; machine-based classifiers label topics/intents without human review (“eyes-off”).
- Taxonomy: topics adapted from Microsoft systems and third-party web-search taxonomies (≈30 topics, see paper’s Table 1); intents derived from user research (11 intents; see Table 2).
- Labeling: each conversation independently assigned a topic and an intent; short/uncertain exchanges labeled “No Topic/No Intent”.
- Analysis approach: ranking topic/intents by frequency, hourly and daily averages, month-to-month delta ranks to identify temporal and modal dynamics.
- Limitations noted by authors: classifier uncertainty for short/ambiguous chats, inability of log-level classification to capture full human-AI relational complexity, and exclusion of enterprise traffic (so professional enterprise usage not represented).
Implications for AI Economics
- Consumer adoption beyond workplace productivity: broad and persistent mobile use—especially for health/advice—signals large consumer demand channels distinct from traditional workplace productivity gains; monetization strategies and product-market fit should reflect this bifurcation.
- Market segmentation by device/context: developers and firms can (and should) design differentiated products, pricing, and features for desktop (workflow augmentation, integrations, enterprise bridging) versus mobile (advice, wellbeing, conversational UX).
- New service markets and value capture: sustained mobile demand for health/advice and wellbeing may expand markets for AI-enabled personal care services, telehealth triage, mental health support, and subscription models tied to daily/24‑hour engagement.
- Labor and productivity effects: desktop augmentation during business hours suggests continued complementarity with knowledge work (task acceleration, editing, coding help), raising questions about task reallocation, skill mix, and measurement of productivity—GDP and labor statistics may undercount AI-mediated task changes if not instrumented for temporal/contextual usage.
- Demand elasticity and time-of-day economics: strong time-of-day and day-of-week patterns imply that pricing, capacity planning, and ad targeting can be temporally differentiated (e.g., peak desktop work hours vs. persistent mobile demand).
- Platform competition and specialization: device-conditioned usage creates niches for specialized agents (health-focused mobile agents vs. productivity-focused desktop agents), increasing incentives for vertical specialization and partnerships (e.g., with healthcare providers, education platforms).
- Safety, regulation, and liability: the high prevalence of advice-seeking (health, relationships, personal growth) on consumer devices raises regulatory concerns—accuracy, duty of care, disclosure, and oversight of non-clinical versus clinical advice; economic modeling should account for compliance costs and risk management.
- Inequality and access: differential patterns by device and time suggest uneven benefits—workers who can leverage desktop-based augmentation in work hours may capture productivity gains more readily than others; public policy may need to consider distributional impacts.
- Measurement and statistics: macroeconomic measurements of AI’s impact should incorporate temporal and modal usage (consumer welfare, nonmarket benefits from 24/7 wellbeing support), not just counts of enterprise deployments or hours saved.
- Product design and externalities: contextualized agents (UI/personality/capability tuned to device/time) can increase engagement and utility but may amplify attention or behavioral externalities (e.g., increased health anxiety or overreliance), which have wider economic and social costs to model.
If you want, I can: - Extract the top topic-intent pairs and hourly trend figures from the paper into a concise table for economic modeling. - Draft brief policy recommendations (regulatory and measurement) oriented to national statistical agencies or competition authorities.
Assessment
Claims (9)
| Claim | Direction | Outcome | Confidence & Evidence | Details |
|---|---|---|---|---|
| We analyze 37.5 million deidentified conversations with Microsoft's Copilot between January and September 2025. Other | null_result | dataset size / analytic sample |
Reading fidelity
high
Study strength
high
|
n=37500000
|
| How people use AI depends fundamentally on context and device type. Task Allocation | positive | variation in usage patterns by device/context |
Reading fidelity
high
Study strength
medium
|
n=37500000
|
| On mobile, health is the dominant topic, which is consistent across every hour and every month we observed - with users seeking not just information but also advice. Task Allocation | positive | topic prevalence on mobile (health dominance) and request type (information vs advice) |
Reading fidelity
high
Study strength
medium
|
n=37500000
|
| On desktop, the pattern is strikingly different: work and technology dominate during business hours, with 'Work and Career' overtaking 'Technology' as the top topic precisely between 8 a.m. and 5 p.m. Task Allocation | positive | topic prevalence on desktop during business hours ('Work and Career' vs 'Technology') |
Reading fidelity
high
Study strength
medium
|
n=37500000
|
| Programming queries spike on weekdays while gaming rises on weekends. Task Allocation | positive | topic frequency by day-of-week (programming vs gaming) |
Reading fidelity
high
Study strength
medium
|
n=37500000
|
| Philosophical questions climb during late-night hours. Task Allocation | positive | hourly topic frequency (philosophical questions) |
Reading fidelity
high
Study strength
medium
|
n=37500000
|
| Relationship conversations surge on Valentine's Day. Task Allocation | positive | date-specific topic frequency (relationships on Valentine's Day) |
Reading fidelity
high
Study strength
medium
|
n=37500000
|
| These patterns suggest that users have rapidly integrated AI into the full texture of their lives, as a work aid at their desks and a companion on their phones. Task Allocation | positive | interpretation of integration/adoption patterns across contexts |
Reading fidelity
high
Study strength
speculative
|
n=37500000
|
| Unlike prior analyses of AI usage, we focus not just on what people do with AI, but on how and when they do it. Other | null_result | research focus (temporal and contextual analysis vs prior content-only analyses) |
Reading fidelity
high
Study strength
medium
|
n=37500000
|