Home Explorer Features Capabilities Platforms Rankings Compare

Click to Do

Click to Do is Microsoft's AI-powered contextual action system on Copilot+ devices that suggests relevant actions based on selected content using on-device NPU processing.

Why It Exists

Users need to quickly act on selected content without copying to other apps or manually choosing between tools.

How It Works

Click to Do is Microsoft's AI-powered contextual action system for Copilot+ Windows devices. Using the device's NPU, the system analyzes selected text, images, or content and suggests relevant actions such as searching, translating, summarizing, or creating content. The system works across any application by detecting the selected content type and context, then providing intelligent, one-click actions. Processing is primarily on-device with optional cloud enhancement.

Everyday Use Cases

  • Content action suggestions
  • Text summarization
  • Image search
  • Translation
  • Content creation
  • Data extraction

User Workflow

  • User selects text, image, or content in any app
  • AI analyzes the selected content and context
  • Relevant actions are suggested in a floating menu
  • User clicks a suggested action
  • Action is executed immediately

AI Processing Flow

Selected content is analyzed by on-device NPU models for type and intent. Text processing handles summarization and extraction locally. Image analysis runs on-device for search and editing. Complex queries may use cloud models. All analysis happens with minimal user data exposure and clear consent prompts.

Inputs / Outputs

Inputs:

  • Selected text
  • Selected images
  • App context
  • Content type
  • User selection
  • System context

Outputs:

  • Action suggestions menu
  • Summarized text
  • Search results
  • Translated content
  • Created content
  • Extracted data

Known Limitations

  • Limited to supported content types
  • Cloud actions require internet
  • Some apps may not support selection detection
  • Suggestion relevance improves over time

Unsupported Scenarios

  • Multilingual translation
  • Handwriting recognition
  • Diagram creation
  • Data visualization
  • Code generation

Performance Notes

Content analysis in under 500ms; On-device actions in under 1 second; Cloud actions 3-7 seconds; NPU usage under 5 percent per active session; Supports 20+ content types

Available On

Shows where this feature is available and how its AI processing works on each platform. Availability may vary by device.

PlatformExecutionOfflineCloudOS / Software
Copilot+ Local Yes No - on-device Windows 11 24H2

Research Status

Confidence and verification reflect how complete documentation is. Fields may show Not assessed or Not yet verified while research is ongoing - this flags gaps, not product deficiencies.

Research StatusVerified
ConfidenceHigh
First introduced2024-09-04
Last updated2026-08-16
Last verified2026-08-16

Research Notes

Available on Copilot+ PCs with Windows 11; Requires Snapdragon X or Intel Lunar Lake NPU; Introduced with Windows 11 2024 Update; Microsoft Copilot ecosystem

Details

VendorMicrosoft
AI PlatformCopilot+
CategoryVision

Capability Mapping

This feature maps to canonical capability family: Visual Understanding.

Canonical capability: Visual Understanding.

  • Visual Understanding - Interprets what is in view and provides contextual understanding or actions.

Grounded in 1 device(s) on Copilot+.

Related Features

This page describes platform-level capability; not device-level support for any specific product.

← Back to Home