Merlin AI is often installed casually, usually as a browser extension someone wants to try for summarising an article, understanding a YouTube video or speeding up an email. At first glance, it feels like a practical addition: lightweight, easy to access and available without repeatedly switching between browser tabs.
For many users, that initial impression is positive. Merlin AI does not demand a complicated setup or a long learning process. It integrates into everyday browsing and writing tasks, which is exactly what makes it appealing.
However, as usage deepens, especially for paying users and people who regularly produce long-form content—some limitations become more noticeable. These are not limited to response length. Query deductions, fair-use restrictions, changing model behaviour and privacy considerations also affect whether Merlin works as a primary AI platform or is better kept as a browser-based assistant
Merlin AI is designed as a browser-based AI assistant that works directly inside websites and web applications. Instead of operating only as a standalone chatbot, it places AI assistance over existing workflows such as reading articles, drafting emails, analysing files and preparing social media content.
Users can open the extension using Ctrl + M on Windows or Command + M on Mac. Merlin also places dedicated AI controls inside selected platforms such as Gmail, LinkedIn, YouTube and X. Its current product pages promote access to several model families, webpage and document chat, image generation, research tools, custom bots, Projects and more than 20 visualisation formats.
In theory, this approach is efficient. Users can select text, ask a question about the current page and continue working without manually copying everything into another application. For short and clearly defined tasks, this remains Merlin’s strongest advantage.
The Chrome Web Store currently shows Merlin with a 4.8/5 rating from approximately 8,800 ratings and around 900,000 users. Google also labels the publisher as following recommended extension practices and having a good record without a history of violations. These are meaningful trust signals, although they do not prove that every paid workflow will perform equally well.
In practice, experienced users, particularly those familiar with the official ChatGPT, Claude or Gemini platforms—sometimes report a gap between expectation and output. Complaints include responses that feel shorter, instructions being missed during longer tasks and a workflow that has become more crowded as additional features have been added.
These experiences are not universal. G2 reviewers also praise Merlin for its browser integration, access to multiple language models and ability to summarise webpages and videos without switching applications. Merlin currently holds 4.1/5 on G2, although that score is based on a relatively small sample of 13 reviews
Based on both general usage and specific user complaints, Merlin AI performs unevenly depending on task complexity:
| Area | Observed or documented performance | Evidence level |
|---|---|---|
| Ease of entry | Easy to install and available from almost any webpage | Verified through the current extension workflow |
| Best use case | Short summaries, quick rewrites, email assistance and questions about the open page | Supported by product features and recurring user feedback |
| Model access | Several AI model families are available through one account | Verified through Merlin’s current product pages |
| Output length | Some paid users report responses that feel shorter than expected | User-reported; requires controlled comparison |
| Context retention | Some users report that detailed instructions are lost during longer tasks | User-reported; not consistently reproduced |
| Query usage | Different models and features deduct different numbers of queries | Confirmed in Merlin’s Query Standards |
| Speed | Short browser-based tasks are commonly described as fast | Supported by G2 and Chrome feedback |
| Long-form reliability | More supervision may be needed for detailed professional work | Reasonable caution based on reported limitations |
| Privacy | The extension may collect URLs and other browsing activity while enabled | Confirmed in Merlin’s Privacy Policy |
One of the most common complaints from paying users is that Merlin’s responses sometimes feel artificially shortened. Users who upgraded expecting output comparable to an official ChatGPT or Claude subscription have reported receiving less depth or needing several prompts to complete a long task.
This becomes particularly frustrating for tasks such as:
For these use cases, repeated prompting can reduce Merlin’s main advantage. A browser assistant is supposed to remove steps, but that benefit becomes smaller when the user has to repeatedly ask the tool to continue, restore missing instructions or expand incomplete sections.
It is important not to describe this only as a fixed “token limit,” because Merlin does not publish one simple output cap that applies to every model and task. The final response can be affected by the selected model, context size, query type, web access, uploaded material and Merlin’s own integration settings.
Merlin advertises up to 102 free queries per day, but one query does not always equal one AI response.
Its published Query Standards assign different deductions to different models. For example, GPT-4o Mini and Claude 3 Haiku are listed at one query per use, while GPT-4o costs 15, Claude 3.5 Sonnet costs 25 and Claude 3 Opus costs 50. Image models can consume 100 or more queries, and enabling live search can double the normal deduction.
| Model or mode listed by Merlin | Published deduction per use |
| GPT-4o Mini | 1 query |
| GPT-4o | 15 queries |
| Gemini 1.5 Flash | 1 query |
| Gemini 1.5 Pro | 15 queries |
| Claude 3 Haiku | 1 query |
| Claude 3.5 Sonnet | 25 queries |
| Claude 3 Opus | 50 queries |
| DALL-E 3 / Render Works v3 | 100 queries |
| Live search | Double the normal model deduction |
The cost of webpage summaries, blog summaries and YouTube summaries can also vary according to the length and language of the source material. Merlin says that its automatic mode chooses a model and deducts queries according to whichever model and mode were selected.
This means the advertised free allowance should not be interpreted as 102 premium-model prompts every day. A casual user choosing lightweight models may complete many small tasks, while a user selecting premium models or web search can use the allowance much faster.
There is also a documentation gap worth noting. Merlin’s Query Standards still list older model versions, while the current login and product pages promote newer choices such as GPT-5, Claude Sonnet 4.5, Grok 4, DeepSeek R1 and Gemini 2.5 Pro. Users should therefore check the live model selector and visible deduction inside their account rather than relying only on the public query table.
One of the more frustrating issues reported by users involves context instability. In practical terms, this means Merlin may not consistently follow every instruction during a long or complicated task.
For example, a user may ask the system to retain a fixed structure, avoid translating a particular section or preserve a specified writing style. The response may begin correctly and then gradually move away from those requirements.
This pattern creates what can be described as instruction fatigue. Instead of accelerating the workflow, the tool requires the user to repeatedly restate the same rules.
The impact becomes more noticeable in workflows that depend on precision:
These problems should not be presented as affecting every user or every model. They need to be tested using a controlled sequence of follow-up prompts.
A useful context-retention test would begin with five fixed instructions and then continue for at least five messages. The reviewer should record exactly which requirement is first ignored and whether selecting a different model changes the result.
Merlin AI works best as a convenience-focused browser assistant for users who want quick help with reading, writing and research directly inside a webpage.
Its core value remains easy to understand. A user can open an article, video, email or social media page and ask for assistance without moving the content into another tool. For short tasks, that can genuinely save time.
The overall user picture is mixed rather than entirely negative. Merlin has a 4.8/5 Chrome Web Store score from around 8,800 ratings and a 4.1/5 G2 score from 13 reviews. Trustpilot presents a different view, with a 2.0/5 score from 21 reviews and frequent complaints concerning subscriptions, support and output quality. Trustpilot also notes that the company has not invited reviews there, so its smaller sample may not represent the wider customer base.
For casual users, students, marketers and professionals completing brief browser-based tasks, Merlin can be useful. The free plan is enough to test its summarisation, rewriting and page-aware features before paying.
Power users should be more cautious. Long-form writers, researchers, analysts and anyone processing many large documents should test output length, context retention and query deductions using their actual workflow. They should also understand that Merlin’s paid access remains subject to documented internal usage limits.
Merlin is therefore not simply a poor version of ChatGPT or Claude. It is a different type of product: an aggregation and productivity layer built around browser convenience.
That convenience is valuable when the task is short and the page already contains the necessary context. It becomes less valuable when accuracy, predictable long-form output, confidential data handling or complete control over the original model experience matters more.
Share your thoughts about this article.
Be the first to post a comment!