On this page+
01 / AI
The logo is not the answer
Four chat windows. Four brands. One slightly tired human asking, ‘Which one should I use?’ The internet usually answers with a champion belt, a benchmark screenshot, or somebody declaring that one model has ‘won’—as if a student researching a presentation and a developer fixing a codebase were entering the same boxing match.
They are not identical products, even when they can all write an email, explain SEO or turn a messy paragraph into something presentable. ChatGPT is a broad assistant and workbench; Gemini is deeply connected to Google’s ecosystem and research tools; Claude leans into careful writing, long-context work, coding and reusable artifacts; Perplexity is built around search-first answers, source discovery and citation trails. Those are tendencies, not iron laws.
“The best AI tool is usually the one that makes the next useful action obvious.” — A less glamorous, more reliable buying rule

For a general all-round assistant, start with ChatGPT. For Google Search, Gmail, Drive and NotebookLM-heavy work, test Gemini. For polished long-form writing, large documents, coding and artifacts, test Claude. For current research where visible citations matter first, start with Perplexity. Then run the same prompt yourself—because your language, plan and task can change the winner.
A small warning before the comparison gets too confident: the app name is not the model name. Plans, regions, usage limits, web access, selected models and feature rollouts change. A review written six months ago can be perfectly sincere and still be stale. AI comparisons age like cut fruit in a Hong Kong office pantry.
02 / AI
What each tool is trying to be
| Tool | Centre of gravity | Where it often feels strongest | Where you should slow down |
|---|---|---|---|
| ChatGPT | General-purpose assistant and workbench | Mixed tasks: drafting, analysis, coding, files, images, voice, projects and research workflows. | Feature and model choice can vary by plan; a fluent answer still needs checking. |
| Gemini | Google-connected multimodal assistant | Google Search-backed research, Workspace context, Canvas, NotebookLM and visual or structured explanations. | Source selection, account permissions and feature availability matter; Google-connected does not mean automatically correct. |
| Claude | Careful collaborator for writing, reasoning, code and artifacts | Long documents, rewriting, code review, structured thinking and self-contained shareable outputs. | A polished tone can make uncertainty look smaller than it is; web search must be enabled or available in the current plan. |
| Perplexity | Search-first research interface | Fast source discovery, direct links, follow-up research, selectable modes and visible citations. | Citations show where claims came from, not that every synthesis is right; it may be less natural for final brand voice. |
The product descriptions are grounded in current official documentation: OpenAI presents ChatGPT as a broad set of capabilities (overview and capabilities help); Google documents Gemini Deep Research, Canvas and Gems (Deep Research, Canvas, Gems); Anthropic documents web search and Artifacts (web search, Artifacts); Perplexity describes Pro Search as multi-source research with direct citations (Pro Search).

Think of them as four colleagues. ChatGPT is the adaptable generalist who can move from spreadsheet to draft to code. Gemini is the colleague already sitting inside Google’s building. Claude brings a red pen, a large document folder and a preference for clean structure. Perplexity walks in holding receipts. None should be left alone with an important fact and no supervision.
03 / AI
The same-prompt test: how to compare them without cheating
Here is the exact prompt used as the editorial benchmark in this article:
You are advising a Hong Kong student or working professional who wants to learn practical AI tools for research and marketing. Compare ChatGPT, Gemini, Claude, and Perplexity for this exact job: explain the difference between SEO and SEM for a Hong Kong business; recommend a realistic 30-day learning plan; cite current official sources; identify what may be outdated or uncertain; and finish with one clear recommendation for a beginner. Use plain English, a compact comparison table, and separate facts from your own judgement. Do not invent prices, rankings, or product features.
This prompt is deliberately ordinary. It asks for research, explanation, a learning plan, current citations and uncertainty—the sort of mixed job a student or working professional actually hands to an AI assistant. It also contains a trap: ‘SEO versus SEM’ is simple enough to explain badly, and current product information is easy to hallucinate if the system is not browsing.
| Control | Keep constant | Why it matters |
|---|---|---|
| Time and date | Run the four tests in the same week and save the exact model, plan and timestamp. | Product features and web results move. A result without a date is half a result. |
| Prompt | Paste the same text, punctuation included; do not improve one tool’s prompt after seeing another’s answer. | Prompt edits quietly turn a comparison into a preference test. |
| Source rule | Ask every tool for current official sources and open the links rather than counting citations. | A citation can be present, broken, secondary or unrelated to the sentence beside it. |
| Context | Use the same language, no extra files, same follow-up policy and the same browsing condition where possible. | A file upload or a connected Drive can change the task completely. |
| Scoring | Score accuracy, source quality, usefulness, uncertainty, Hong Kong relevance and edit time. | ‘I liked the tone’ is useful feedback, but it is not the whole test. |
Do not compare Perplexity with citations switched on against another tool with browsing switched off and then call the result a model verdict. That is not a fair fight; it is comparing a librarian with a person who was told to keep their eyes closed. If the products cannot be made feature-equivalent, record the difference as part of the product experience.
This is a practical product comparison, not a controlled scientific benchmark. Model routing, plan limits, regional rollout, system instructions and live web results can change the output. Use the test as a repeatable method, not as a permanent league table.
04 / AI
What the four answers tend to feel like
| Test dimension | ChatGPT | Gemini | Claude | Perplexity |
|---|---|---|---|---|
| Explaining SEO and SEM | Usually balanced and adaptable; can shift from beginner explanation to technical detail quickly. | Often good at structured explanation and current-search framing, especially when Search is used. | Often clear, measured and easy to read; may add useful caveats without making the answer feel like a legal notice. | Usually concise and source-forward; often shows where the explanation came from before polishing the prose. |
| 30-day learning plan | Good at turning a vague goal into steps, checklists and a portfolio task. | Useful when the plan includes Google tools, Workspace or research sources. | Good at sequencing and tone; often turns the plan into a clean, teachable document. | Useful for finding course and tool references, but the user may need to reshape the research into a personal routine. |
| Current official sources | Can search or research when enabled; still open the links and verify dates. | Deep Research uses Google Search by default and can include selected sources such as Drive or Gmail when available. | Web search can ground current answers when enabled; do not assume it is always active. | Citations and source discovery are the product’s centre stage, especially in Pro Search. |
| Final writing polish | Flexible: casual, professional, persuasive, bilingual or highly formatted. | Strong for visual, structured and Google-workflow outputs; style varies by prompt. | Often the smoothest for nuanced rewriting and long-form coherence. | Good research skeleton; final brand voice may need more hands-on editing. |
| Main risk | Confident synthesis can blur what is known, inferred or generated. | Connected sources can create a false sense that the answer is automatically verified. | Elegant language can hide a wrong assumption if you do not inspect sources. | A long citation list can look like proof even when source-to-claim matching is weak. |

These are tendencies, not report-card grades. A strong prompt can make a quiet tool sing, and a sloppy prompt can make a powerful tool produce soup. The biggest difference in everyday work is often not raw intelligence; it is how easily the product helps you gather evidence, keep context, edit the result and take the next step.
For SEO work, this is especially important. Ask an AI tool for ten keywords and you may receive ten plausible-looking phrases. Ask it to map search intent, identify the source for a claim, mark assumptions and propose a test you can run in Search Console, and suddenly you are measuring judgement rather than autocomplete.
05 / AI
The actual trade-offs: sources, long context, files and workflow
Need receipts
Perplexity’s source-forward design is convenient when you want links visible early and often.
Need a clean rewrite
Claude is often a comfortable second pair of eyes for long, awkward drafts and structured documents.
Need a mixed toolbox
ChatGPT is a strong starting point when one task jumps between files, analysis, writing, voice or code.
Need Google context
Gemini is the obvious test when your work already lives in Search, Drive, Gmail or NotebookLM.
Need a repeatable routine
The best product is the one whose limits and context rules you understand well enough to use every week.
| Workflow question | What to inspect before choosing |
|---|---|
| Can it browse? | Is web access automatic, optional, plan-limited or disabled in your current account? |
| Can it cite? | Are citations direct and relevant, and can you open the source behind each important claim? |
| Can it handle my files? | Check file size, formats, privacy settings, context limits and whether the answer actually used the file. |
| Can I reuse the result? | Look for Projects, Canvas, Artifacts, custom instructions, Gems, export, sharing and version control. |
| Can my team use it safely? | Review account type, training/data controls, connectors, retention, permissions and confidential-client rules. |
| Can it work in Hong Kong language? | Test Traditional Chinese, Cantonese wording, English source terms and whether it keeps terms consistent instead of translating blindly. |
Google’s current Gemini documentation says Deep Research uses Google Search by default and can add sources such as Gmail, Drive, uploaded files or NotebookLM when those connections are available. Perplexity’s current help documentation describes selected premium sources and direct citations. Anthropic’s help page explains that web search must be enabled through the current interface or workspace settings. The word ‘can’ is doing real work in all three descriptions—availability depends on account, plan and rollout.
Do not paste a client’s confidential traffic report, customer list, unreleased product plan or personal data into four consumer accounts just to compare tone. An AI answer that saves twenty minutes is not a bargain if it creates a data problem you cannot explain.
One more Hong Kong-specific nuisance: language switching. A tool may answer in clean written Chinese while missing the Cantonese search intent behind a phrase, or it may translate an English SEO term into something nobody in the market actually searches. Always keep the original query, the translated version and a native-reader check in your notes.
06 / AI
Which one should a student or working professional use?
| Your situation | Best first test | Why |
|---|---|---|
| Student researching a report | Perplexity for source discovery, then ChatGPT or Claude for restructuring and explanation. | Separate finding evidence from turning it into a readable argument. |
| Student learning marketing or SEO | ChatGPT or Gemini for guided practice; use Google Search Central and tool documentation as the factual base. | The assistant can coach, but official documentation should anchor the fundamentals. |
| Working professional writing reports | Claude and ChatGPT side by side on the same anonymised draft. | Compare editing quality, instruction-following and how much cleanup remains. |
| Working professional in Google Workspace | Gemini is worth testing first, especially for connected research and document workflows. | The ecosystem connection may save more time than a small difference in prose quality. |
| Marketer doing current research | Perplexity and Gemini/ChatGPT with web research enabled. | Compare source visibility, freshness and the time needed to verify claims. |
| Developer or technical learner | ChatGPT and Claude on the same bug, specification or repository excerpt; Gemini if Google tools are central. | Code quality, context handling and the ability to explain trade-offs matter more than a generic chatbot ranking. |

My practical recommendation is a two-tool habit, not a four-tab religion. Use one tool to discover or challenge the evidence, and another to shape the final work. That separation makes it easier to notice when a beautifully written paragraph has no solid floor underneath it.
If you only want one free starting point, begin with the tool whose interface you will actually open tomorrow. A theoretically superior assistant that you avoid because the workflow feels clumsy is not superior in your life. This is not philosophy; it is calendar management wearing a small AI hat.
07 / AI
A repeatable same-prompt scorecard
| Category | 0 points | 1–3 points | 5 points |
|---|---|---|---|
| Accuracy | Material errors or invented product details. | Mostly right but with unsupported or vague claims. | Claims are accurate, qualified and easy to verify. |
| Source quality | No sources or irrelevant links. | Some relevant links, but source-to-claim matching needs work. | Primary/current sources are attached to the claims they support. |
| Task usefulness | Generic advice that could fit anyone. | Useful outline but still needs major reshaping. | A realistic plan with concrete next actions and constraints. |
| Uncertainty | Speaks as if everything is certain. | Mentions limits but does not locate them clearly. | Separates fact, inference, unknown and what to verify next. |
| Hong Kong relevance | Ignores language, market and local context. | Mentions Hong Kong but remains generic. | Handles HK language, examples, search behaviour and practical constraints naturally. |
| Edit time | You have to rewrite almost everything. | Needs moderate cleanup and fact checking. | Can be used after sensible human review. |
Keep the raw answers. Save the prompt, model name, plan, browsing status, timestamp, citations and your score. Without that record, ‘Claude was better’ or ‘Gemini felt smarter’ is just workplace folklore, one step away from becoming a slide deck.
- First read: does it answer the job?A beautifully written non-answer still fails. Mark whether every requested part appears.
- Second read: can the important claims be checked?Open sources, inspect dates and reject links that merely contain the right words.
- Third read: what did it assume?Look for hidden audience, market, budget, language or product assumptions.
- Fourth read: what can you use tomorrow?Keep the parts that reduce real work: a checklist, brief, table, experiment or decision.
- Final read: where does a human need to sign off?Facts, privacy, client claims, money, health, law and public-facing recommendations need human ownership.
A smooth paragraph is a style result. It is not a source, a calculation or a guarantee. Read the receipts.
08 / AI
FAQ: the blunt answers people usually want
Which is the most accurate: ChatGPT, Gemini, Claude or Perplexity?
There is no universal winner, and ‘accuracy’ changes by task. A tool with current search may be better for today’s product fact; a tool with stronger context handling may be better for your long document; a tool with clearer citations may be easier to audit. Treat accuracy as something you test and verify, not a badge the app owns forever.
Which one is best for SEO work?
Use more than one role. Perplexity or Gemini can help discover current sources and terminology; ChatGPT or Claude can help turn a verified brief into an outline, audit explanation or client-facing draft. None replaces Search Console, Analytics, technical inspection, native-language judgement or your responsibility for claims.
Which is best for students?
Start with the one that helps you understand rather than merely submit. ChatGPT and Gemini can tutor and structure practice; Perplexity can make source discovery visible; Claude can help revise a long draft. Ask for explanations, counterarguments and source links, then write at least part of the work yourself. Your lecturer can usually tell when you did not.
Which is best for working professionals?
Test the tool against a real anonymised workflow: meeting notes to action list, report to executive summary, research to recommendation, or brief to first draft. Measure time saved after fact checking, not time saved before it. Also review your organisation’s data rules before connecting files or pasting client material.
Can I trust an AI answer because it includes citations?
No. Citations improve auditability, not automatic truth. Open the link, check whether it supports the exact claim, inspect the date and see whether the answer quietly stretched a narrow source into a broad conclusion.
Do I need to pay for all four?
Almost certainly not. Test free or existing access first, choose one primary tool and add a second only when it solves a specific gap—such as source discovery, Google-connected work, long-document editing or a different coding workflow. Four subscriptions can become a very expensive way to avoid choosing a workflow.
Run the same prompt, keep the raw outputs, score evidence and edit time, and choose the tool that helps you make better decisions—not the one that wins the loudest online argument.
ChatGPT, Gemini, Claude and Perplexity are not four identical doors leading to one room. They are four different corridors, each with useful shortcuts and a few suspicious-looking doors. Walk the one that matches the work you actually do, and keep your own judgement switched on.
RESEARCH + CREDITS
Sources, product notes and image credits
- OpenAI ChatGPT overview
- OpenAI ChatGPT capabilities
- OpenAI ChatGPT Pro help
- Google Gemini 3.1 Pro
- Google Gemini Deep Research
- Google Gemini Canvas
- Google Gemini Gems
- Anthropic Claude Opus 5
- Anthropic Claude Sonnet 4.6
- Anthropic web search help
- Anthropic Artifacts help
- Perplexity Pro Search
- Perplexity premium data sources
- Perplexity
Photo credits
- 01-hku-lecture.jpg — real classroom/workshop photograph from www.thestandard.com.hk. Direct image: original file.
- 02-uowhk-workshop.jpg — real classroom/workshop photograph from www.uowchk.edu.hk. Direct image: original file.
- 03-polyu-hkcc-seminar.jpg — real classroom/workshop photograph from www.hkcc-polyu.edu.hk. Direct image: original file.
- 04-digital-marketing-class.webp — real classroom/workshop photograph from ingoacademy.ntu.edu.tw. Direct image: original file.
Research current through 2026-09-16 UTC. AI model names, plans, limits, web access, source connectors and product features can change; the comparison records public product documentation and a repeatable same-prompt method, not a permanent ranking. Image reuse rights remain with the original publishers or photographers. Product features, model routing, plans and availability can change. Verify the live product page before making a purchase or operational decision.