Click to Do
Click to Do is Microsoft's AI-powered contextual action system on Copilot+ devices that suggests relevant actions based on selected content using on-device NPU processing.
Why It Exists
Users need to quickly act on selected content without copying to other apps or manually choosing between tools.
How It Works
Click to Do is Microsoft's AI-powered contextual action system for Copilot+ Windows devices. Using the device's NPU, the system analyzes selected text, images, or content and suggests relevant actions such as searching, translating, summarizing, or creating content. The system works across any application by detecting the selected content type and context, then providing intelligent, one-click actions. Processing is primarily on-device with optional cloud enhancement.
Everyday Use Cases
- Content action suggestions
- Text summarization
- Image search
- Translation
- Content creation
- Data extraction
User Workflow
- User selects text, image, or content in any app
- AI analyzes the selected content and context
- Relevant actions are suggested in a floating menu
- User clicks a suggested action
- Action is executed immediately
AI Processing Flow
Selected content is analyzed by on-device NPU models for type and intent. Text processing handles summarization and extraction locally. Image analysis runs on-device for search and editing. Complex queries may use cloud models. All analysis happens with minimal user data exposure and clear consent prompts.
Inputs / Outputs
Inputs:
- Selected text
- Selected images
- App context
- Content type
- User selection
- System context
Outputs:
- Action suggestions menu
- Summarized text
- Search results
- Translated content
- Created content
- Extracted data
Known Limitations
- Limited to supported content types
- Cloud actions require internet
- Some apps may not support selection detection
- Suggestion relevance improves over time
Unsupported Scenarios
- Multilingual translation
- Handwriting recognition
- Diagram creation
- Data visualization
- Code generation
Performance Notes
Content analysis in under 500ms; On-device actions in under 1 second; Cloud actions 3-7 seconds; NPU usage under 5 percent per active session; Supports 20+ content types
Available On
Shows where this feature is available and how its AI processing works on each platform. Availability may vary by device.
| Platform | Execution | Offline | Cloud | OS / Software |
|---|---|---|---|---|
| Copilot+ | Local | Yes | No - on-device | Windows 11 24H2 |
Research Status
Confidence and verification reflect how complete documentation is. Fields may show Not assessed or Not yet verified while research is ongoing - this flags gaps, not product deficiencies.
| Research Status | Verified |
| Confidence | High |
| First introduced | 2024-09-04 |
| Last updated | 2026-08-16 |
| Last verified | 2026-08-16 |
Research Notes
Available on Copilot+ PCs with Windows 11; Requires Snapdragon X or Intel Lunar Lake NPU; Introduced with Windows 11 2024 Update; Microsoft Copilot ecosystem
Details
| Vendor | Microsoft |
| AI Platform | Copilot+ |
| Category | Vision |
Capability Mapping
This feature maps to canonical capability family: Visual Understanding.
Canonical capability: Visual Understanding.
- Visual Understanding - Interprets what is in view and provides contextual understanding or actions.
Grounded in 1 device(s) on Copilot+.
Related Features
- Circle to Search - Cross-platform equivalent
- Visual Intelligence - Complementary
This page describes platform-level capability; not device-level support for any specific product.