Gemini and NotebookLM, ChatGPT, Claude, and Microsoft Copilot compared across pricing, memory, projects, reusable skills, agents, scheduling, connectors, design, images, input constraints, platform surfaces, and organizational boundaries.
Confirmed factComparative analysisEntitlement caveatForecastResearch date: July 20, 2026
Decision brief
Choose the ecosystem around the work product
The platforms have converged on chat, memory, files, research, images, and agentic behavior. Their economic value now depends more on the surrounding workflow than on a single model benchmark.
01
No single plan dominates every workload.
Google is the strongest low-price ecosystem bundle. ChatGPT is the broadest general assistant. Claude is the most coherent personal workflow-and-design workbench. Microsoft is strongest when the final artifact must remain inside Office.
Lowest standalone priceGoogle AI Plus at $4.99, followed by ChatGPT Go at $8.
Best personal composition systemClaude Pro: Projects + Skills + MCP + Code + Cowork + Design.
Best embedded productivity valueMicrosoft 365 Personal when Office and 1 TB storage are already required.
Pricing and entitlement documentation changes faster than support pages.
Google’s US plan page currently surfaces AI Plus at $4.99 with 400 GB. OpenAI’s Go help page includes Projects, Tasks, custom GPTs, and Library, while the pricing matrix emphasizes those under Plus. Verify the authenticated checkout and account UI before purchasing.
Price floor
The lowest personal paid tiers
Monthly equivalents use published annual prices where available. They do not normalize storage, Office applications, family benefits, promotional pricing, taxes, or regional availability.
A Project, Notebook, GPT, Gem, Skill, Agent, and Copilot Page are not interchangeable. The matrix describes practical availability on the lowest personal paid tier.
03
Capability
Google AI Plus
ChatGPT Go
Claude Pro
M365 Personal
Persistent workspaces
Gemini Notebooks + NotebookLM
Projects
Unlimited Projects
Copilot Notebooks / Pages vary by surface
Cross-chat memory
Personalization + activity context
Saved memory + chat history
Memory + past-chat search
Saved memory + chat-history inference
Reusable assistant
Gems
Custom GPTs
Skills + Projects
Agents and Notebook instructions
Portable skill format
Vendor-specific
Skills rollout is entitlement-dependent
Open Agent Skills format
Vendor-specific
Scheduled work
Scheduled actions
Tasks
Remote Cowork schedules
Fragmented; 25 agent tasks ≠ universal scheduling
Custom external tools
Google integrations; advanced agent surfaces higher-tier
Apps; custom MCP stronger on organizational tiers
Remote MCP + local extensions
Native Office; broader custom agents favor business
Native raster images
Strong
Strong
Structured visuals, not conventional photo generation
Designer/Create with credits
Editable design canvas
Canvas
Canvas / Work surfaces
Claude Design
Pages + Office canvases
Terminal / IDE workflow
Repository ingestion and coding surfaces
Expanded Codex is a Plus advantage
Claude Code included
Strong Office; developer workflow less central
Desktop-native advantage
Web-first; mobile Android integration is material
Work and local integrations expand on desktop
Code, Cowork, Design, local extensions
Full desktop Office applications
Memory model
Memory has four distinct layers
Do not treat “memory” as one feature. A platform may remember user preferences while failing to retrieve a prior project conversation or a source-grounded fact.
04
ProfileStable preferences and user facts.
HistoryRetrieval or inference from previous conversations.
WorkspaceProject- or notebook-scoped context.
KnowledgeUploaded sources and canonical reference files.
ExecutionState retained by tasks, agents, or tools.
ChatGPT
Separates saved memories, chat-history inference, Project memory, GPT instructions, and Temporary Chat. Mature personalization, but scope can be misunderstood when global and project context coexist.
Claude
Combines memory entries, past-chat search with citations, project summaries, project-specific context, incognito chats, and import/export. It is currently the most inspectable personal memory model.
Gemini
Personalizes from prior conversations and connected Google services. Notebooks maintain source-and-instruction continuity. Gems and Live may not inherit every personalization behavior.
Microsoft Copilot
Separates saved memories, chat-history inference, custom instructions, and the contextual data available inside each Microsoft 365 application. Standalone Copilot and embedded Copilot are different surfaces.
Architectural recommendation
Memory should accelerate retrieval, not become the system of record. Preserve decisions, prompts, sources, design tokens, and procedures in versioned portable files.
Projects, skills, GPTs, Gems, agents
Claude has the cleanest personal composition model
Its primitives divide responsibilities more explicitly: Projects hold context, Skills encode reusable behavior, MCP supplies tools and data, Cowork executes workflows, Design creates visual artifacts, and Code operates on software.
05
ChatGPT
Projects for persistent work
Custom GPTs for packaged assistants
Tasks for scheduling
Apps for external data and actions
Managed workspace agents favor Business and Enterprise
Claude
Projects and project memory
Custom Skills and plugins
Remote MCP and local extensions
Cowork tasks and schedules
Code and Design in the same subscription
Gemini
Gems as reusable assistants
Gemini Notebooks synchronized with NotebookLM
Canvas for documents and applications
Scheduled actions
Spark and advanced workflows favor higher tiers
Microsoft
Copilot Notebooks and Pages
Application-specific agents
Agent and editing modes inside Office
25 monthly agent tasks on Personal
Researcher and Analyst require Premium
Do not compare feature names without testing their operational boundary.
“Agent” may mean a research mode, an Office document generator, a scheduled browser task, a managed workspace service, or a reusable assistant with no autonomous execution.
Scheduling
Personal scheduled execution is becoming standard
The meaningful differentiator is whether work runs remotely, whether it can use connectors, and whether approvals and failures are observable.
06
Platform
Mechanism
Remote execution
Key limitation
ChatGPT Go
Tasks / scheduled prompts
Yes, for supported task types
Tool availability and quotas vary by task and plan.
Google AI Plus
Scheduled actions
Yes
Connected-service and feature availability varies by account, region, and surface.
Claude Pro
Cowork scheduled tasks
Yes; computer need not remain awake
Shared plan usage and connector security require attention.
M365 Personal
Agent tasks and product-specific automation
Depends on agent and application
25 agent tasks do not imply a general-purpose personal scheduler; advanced agents require Premium.
Input and retrieval limits
Nominal upload size is not effective comprehension
A large file can fit while important details remain unretrieved. Visual extraction, context allocation, page count, source format, and retrieval strategy matter more than raw megabytes.
07
Platform
Published chat limits
Visual behavior
Practical concern
ChatGPT
512 MB/file; text documents capped at 2M tokens; images 20 MB; spreadsheets ≈50 MB
PDF visual retrieval depends on plan and surface
A visually rich PDF can be less understood than the file-size limit suggests.
Claude
500 MB/file; up to 20 files/chat; project file limit 30 MB
Text and visuals in PDFs under 100 pages; non-PDF documents generally text-only
Split large visual documents; Project and chat upload constraints differ.
Gemini
Up to 10 files/prompt; most files 100 MB; video 2 GB; code repository up to 5,000 files/100 MB
Strong image, audio, video, and repository ingestion
Large-source fidelity improves on higher tiers with larger context and limits.
Microsoft Copilot
Up to 20 files/conversation; 50 MB/file
Supports Office, PDF, image, text, and data formats
Long whole-document operations may omit content; localized Office edits are a stronger fit.
Retrieval validation pattern
Plant unique canary facts at 5%, 50%, and 95% of a source. Add one inside a table and one inside an image. Measure answer accuracy, omissions, and citation fidelity.
Connectors
Claude is the most open personal tool-composition environment
Google and Microsoft have deeper native ecosystems. Claude has the clearest personal path to vendor-neutral external tools. ChatGPT’s strongest managed custom-agent controls remain organizational.
08
Claude
Remote MCP connectors and local desktop extensions support personal tool composition. Plugins can bundle Skills, connectors, and sub-agents. The same openness increases indirect prompt-injection and authorization risk.
ChatGPT
Apps connect external services and can combine interfaces, search, actions, and data. Managed internal knowledge, custom MCP policy, approval control, and persistent workspace agents are materially stronger in Business and Enterprise.
Gemini
Broad consumer integration with Gmail, Calendar, Drive, Docs, Keep, Tasks, Photos, YouTube, Android, and other services. The integration model is powerful but less portable than MCP.
Microsoft
Native depth in Word, Excel, PowerPoint, Outlook, OneDrive, and Microsoft 365 is the primary advantage. Full organizational grounding and Copilot Studio belong to the business architecture.
Image generation and design
“Best design” depends on the editable target
Raster generation, structured UI design, implementation code, and finished Office deliverables are separate production stages.
09
Gemini
Best fit: images, editing, multimodal references, infographics, video, and creative media systems.
Economic edge: these capabilities enter at the lowest subscription price.
ChatGPT
Best fit: iterative image generation connected to research, writing, coding, Canvas, and broader artifact workflows.
Tier effect: Plus improves complexity, accuracy, and throughput.
Claude
Best fit: editable UI concepts, design systems, SVG, HTML, diagrams, and design-to-code handoff.
Important distinction: structured design rather than conventional photorealistic chat image generation.
Microsoft
Best fit: final Word, Excel, PowerPoint, Designer, Create, Paint, and Photos artifacts.
Personal quota: 60 image credits each month; Premium provides extensive use.
Output
Primary choice
Reason
Secondary choice
Photorealistic or illustrative image
Gemini
Low-cost multimodal and media breadth
ChatGPT
Editable UI concept
Claude Design
Design-system interpretation and implementation handoff
ChatGPT Canvas / Work
Research-grounded presentation draft
Gemini Notebook / NotebookLM
Source-corpus grounding
Microsoft PowerPoint agent
Production HTML or application
Claude Code
Design + Skills + MCP + code continuity
ChatGPT Codex
Finished Office artifact
Microsoft 365
Native format and application controls
Claude file creation
SVG and structured diagrams
Claude
Code-native editable visual output
ChatGPT
Cross-platform design rule
Transport the design system, not the screenshot. Preserve tokens, components, states, responsive rules, accessibility behavior, and reference screens in neutral files.
Desktop and mobile
Desktop unlocks execution; mobile unlocks context capture
The desktop advantage is largest when local applications, files, IDEs, browser automation, or Office formatting are required. Mobile is strongest for voice, camera, screen context, location, and quick continuation.
10
Desktop differentiators
Claude: Code, Cowork, Design, local MCP extensions, browser and Office extensions
Microsoft: full Word, Excel, PowerPoint, and Outlook applications
Claude: conversation continuity, capture, review, and remote Cowork task assignment
Microsoft: conversational Copilot and lightweight Office review/editing
Personal versus business
Governance—not intelligence—is the durable boundary
Personal products can receive experimental features first. Business products add identity, policy, retention, data controls, auditability, approved connectors, and managed distribution.
11
ChatGPT
Personal plans support GPTs, Projects, Tasks, memory, files, and agentic modes. Business and Enterprise add managed workspace agents, company knowledge, custom MCP governance, administration, and organizational deployment.
Claude
Pro exposes unusually broad personal functionality. Team and Enterprise primarily add administration, SSO, role controls, organization-wide Skills, enterprise search, and connector governance.
Google
Google AI consumer plans use personal Google accounts. Workspace Gemini is a separate product and policy surface. A personal feature is not automatically available in a managed Workspace tenant.
Microsoft
Personal and Premium are consumer products. Organizational Graph grounding, Teams and SharePoint context, Copilot Studio, managed agents, and enterprise controls belong to business licensing.
Cost-optimized portfolios
Buy complementary capabilities, not redundant chat
Annual equivalents are used for Claude Pro and Microsoft 365 Personal where noted.
12
Minimal viable assistant
ChatGPT Go for memory, Projects, GPTs, Tasks, files, and images.
$8/mo
Lowest-cost broad ecosystem
Google AI Plus for NotebookLM, Google context, media, and 400 GB storage.
$4.99/mo
Broad two-engine validation stack
Google AI Plus + ChatGPT Go. Strong research, media, memory, assistants, and cross-checking.
$12.99/mo
Builder and media stack
Claude Pro annual + Google AI Plus. Strongest mix of Skills, MCP, Code, Design, Notebook, images, and video.
≈$21.66/mo
Builder and Office stack
Claude Pro annual + Microsoft 365 Personal annual. Architect and build in Claude; finish in native Office.
≈$25/mo
Four-platform evaluation laboratory
All four lowest paid tiers. Appropriate for standards, migration, portability testing, and comparative evaluations.
≈$37.99/mo
Interoperability blueprint
Make every vendor an adapter
The canonical capability package should survive cancellation of any subscription and be deployable into GPTs, Skills, Gems, Notebooks, or Copilot agents.
13
01Canonical knowledge
Markdown, source manifests, glossary, examples, decisions, and provenance.
02Reusable behavior
Skills, workflows, policies, prompts, and deterministic scripts.
03Tools and data
MCP, OpenAPI, least-privilege credentials, approvals, and audit events.
04Vendor adapters
GPT, Claude Skill, Gem, Notebook, or Copilot package.
05Shared evaluations
Run identical test cases across platforms and surfaces.
The highest-risk errors come from inconsistent entitlements, false confidence in long context, hidden memory scope, connector injection, shared quotas, and mismatched personal/business identities.
14
Entitlement drift and contradictory documentation
Capture the authenticated plan matrix, checkout price, model picker, and capability settings at purchase time. Revalidate monthly. Treat a marketing page as evidence, not as a service-level commitment.
Long-context retrieval failure
Test canary facts at multiple positions and representations. Score precision, recall, citation accuracy, omission rate, and sensitivity to distractors.
Memory-scope leakage
Create a unique value in one project. Query it from another project, a new chat, voice, mobile, a custom assistant, and a connected application. Repeat after deleting the source chat and clearing memory.
Indirect prompt injection through connectors
Assume web pages, documents, email, calendar invitations, source-code comments, and MCP tool output contain hostile instructions. Separate read and write capability. Require confirmation for destructive actions. Preserve tool provenance.
Shared quota interference
Measure accepted outputs per dollar and per five-hour window. A heavy coding, video, research, or agent run can consume disproportionately more capacity than ordinary chat.
Cross-surface inconsistency
Run the same evaluation on browser, desktop, iOS, Android, voice, Office or Google application, and IDE. Compare memory, files, citations, formatting, tool access, model selection, and approval behavior.
Personal-to-business migration failure
Validate identity, retention, data residency, connector approval, administrative policy, audit logging, distribution, and deletion—not merely whether the same prompt produces a similar answer.
Security boundary
A connector-enabled personal assistant is not an enterprise agent. The capability may look identical while authorization, logging, retention, and administrative controls differ materially.
What we know and what is likely next
Competition is shifting from models to operating systems for work
Confirmed direction and informed forecasts are separated below.
15
Confirmed direction
All four provide persistent context surfaces
Reusable assistants and agentic execution are converging
Scheduling is entering personal tiers
Artifacts are becoming editable rather than chat-only
Desktop and application integrations are strategic
Business differentiation centers on governance
Forecast: 6–18 months
Skills become a common, partially portable packaging unit
Scheduled prompts and autonomous agents converge
Design-to-code retains more component structure and tokens
Memory becomes more inspectable and auditable
Bundling pressure favors Google and Microsoft economics
Provider-neutral evaluations become essential procurement evidence
Recommended starting strategy
Start with Google AI Plus when price and ecosystem breadth dominate. Start with ChatGPT Go when a general assistant, Projects, GPTs, and Tasks are central. Choose Claude Pro when reusable workflow engineering and UI implementation justify the premium. Treat Microsoft 365 Personal as an Office-suite decision rather than a pure chatbot purchase.
Provenance
First-party source registry
Sources were accessed July 20, 2026. Product entitlements, regional prices, quotas, model names, and previews can change without notice. The report favors official product and support documentation.
16
O1
ChatGPT pricingPlans, model access, image, research, Projects, Tasks, GPTs, Codex, and Work distinctions.
Methodology limitation
This is a product-capability comparison, not a controlled model-quality benchmark. “Best” judgments are explicit analysis based on workflow breadth, editability, portability, integration, and price—not a claim that one model wins every prompt.