Briefly
- Alibaba upgraded Qwen Deep Analysis with one-click webpage and podcast technology.
- In testing, Qwen and Gemini tied for accuracy, each outperforming ChatGPT and Grok
- Total, Qwen received for analysis depth and shareable internet output, whereas Gemini led multimedia high quality
Qwen, the devoted AI analysis group throughout the Chinese language tech big Alibaba, launched a big improve to its AI chatbot final week, enabling customers to generate complete analysis paperwork on any subject.
You may then simply convert these paperwork into clear webpages or multi-speaker podcasts with only a few clicks.
Qwen Chat is just like ChatGPT, DeepSeek, or Claude when it comes to UI and is obtainable worldwide totally free.
Qwen Deep Analysis simply obtained a significant improve. ⚡️
It now creates not solely the report, but additionally a stay webpage 🌐 and a podcast 🎙️ – Powered by Qwen3-Coder, Qwen-Picture, and Qwen3-TTS.
Your insights, now visible and audible. ✨
👉 https://t.co/wESb7vfAnD pic.twitter.com/eRvjKU222O— Qwen (@Alibaba_Qwen) October 21, 2025
The brand new performance runs on three open-source fashions working in live performance: Qwen3-Coder handles internet construction, Qwen-Picture generates inline graphics, and Qwen3-TTS powers dynamic audio narration.
Regardless of counting on open-source fashions, the end-to-end expertise—together with analysis execution, internet deployment, and audio technology—is hosted and operated by Qwen as a managed service.
The workflow begins inside Qwen Chat, the place customers pose analysis questions. The AI conducts internet searches after some clarifications, analyzes knowledge from public sources, and generates a complete report with citations.
From there, two new choices seem: “Net Dev” produces a stay, professional-grade webpage routinely deployed and hosted by Qwen, full with inline graphics.
“Podcast,” in the meantime, affords an audio dialogue that includes dynamic multi-speaker narration, with 17 host voices and 7 co-host choices.

Testing the fashions
To evaluate how Qwen stacked up as a analysis instrument, we ran the identical advanced analysis question throughout it, Gemini, ChatGPT, and Grok. The duty, which may be reviewed on our GitHub repo, was to investigate philosophical and scientific arguments for and towards God’s existence. Every mannequin generated a full analysis report. The analysis concerned 5 standards: accuracy of claims and citations, info offered, readability of rationalization, mental richness, and general high quality.
TL;DR: Qwen Deep Analysis wins for analytical depth, quotation, and its distinctive auto-generated webpages, making it superb for lecturers and creators. It is also the most effective all-in-one free different for researchers. However Gemini nonetheless leads in audio and video high quality, whereas ChatGPT and Grok stay wonderful for informal use however lack Qwen’s attain and Google’s polish.
Here is a extra in-depth overview:
Accuracy: Had been philosophical positions and scientific claims represented appropriately, with correct supply attribution?
Qwen nailed the main points. When discussing the cosmological argument, it correctly cited tutorial sources like Bertrand Russell’s “Why I am not a Christian” and the debate between William Lane Craig and Peter Atkins, with particular references. In contrast to different AI researchers like Perplexity’s or Grok, the vast majority of sources are respected and tutorial, generally even the Unique Supply. It included hyperlinks from Stanford, Princeton, Oxford, Drew, however added pertinent evaluation from Quora and Fb when related.
Gemini matched this precision with 94 numbered citations, a few of which have been duplicated when referenced in several components of the report.
It appropriately distinguished between ideas. Each averted sloppy errors, comparable to conflating biblical literalism with normal theism.
ChatGPT relied closely on the Stanford Encyclopedia of Philosophy, however generally oversimplified. Grok gave correct summaries however with vaguer attribution—saying issues like “traced to Plato, Aristotle” with out particular works.
Outcome: Qwen and Gemini have been the most effective.
Data Offered: How thorough was the analysis?
Qwen was the one mannequin to incorporate a piece referred to as “Critiques of Atheism: The Burden of Proof and the Nature of Proof.” This part examined a kind of debate not one of the others touched. It distinguished between “weak atheism” (skepticism towards God claims) and “gnostic atheism” (constructive assertion God would not exist), and cited particular atheist thinkers like Gary Whittenberger’s “past an inexpensive doubt” customary.

Here is an instance passage from Qwen: “One of the vital contentious points is the burden of proof. Bertrand Russell famously illustrated this together with his teapot analogy: simply as he couldn’t show {that a} tiny teapot doesn’t orbit the solar between Earth and Mars, he argued that theists couldn’t show that God does exist.”
No different mannequin went this deep into burden-of-proof debates as a result of it most likely was not central to the subject. Gemini got here shut with robust protection of consciousness arguments and the “God-of-the-gaps” critique. ChatGPT included pragmatic arguments like Pascal’s Wager and explored real-world implications for ethics and coverage. Grok stored it concise—about one-third the size of Qwen’s report—however added a useful abstract desk.
Outcome: Qwen was probably the most exhaustive.
Readability: How was the analysis expressed?
Grok used a clear desk to prepare arguments by kind (Philosophical vs. Scientific, For vs. Towards). Its part breaks have been specific: “Philosophical Arguments,” “Scientific Arguments,” “Sudden Element.” Anybody may scan it rapidly.
ChatGPT used tons of parenthetical clarifications, making advanced concepts extra digestible. Instance: “if God’s existence is even attainable (i.e., logically coherent), then God exists essentially.” The “(i.e., logically coherent)” helps readers who aren’t philosophy majors.
Qwen and Gemini, then again, have been extra tutorial of their type. Qwen organized the content material beneath formal headings like “Theistic Arguments for God’s Existence: Cosmological and Teleological Foundations,” which made the entire studying really feel very dense, regardless of its accuracy. Gemini used Roman numerals (I. Introduction, II. Philosophical Arguments), which seemed structured however required nearer studying.
Each Qwen and Gemini goal researchers doing severe work. ChatGPT and Grok goal broader audiences.
Outcome: ChatGPT introduced info probably the most clearly, adopted by Grok.
Range of sources: Does the analysis draw from diversified traditions, disciplines, and views?
Qwen built-in technical philosophy (kalām, PSR, modal S5 logic) with stay scientific debates (Massive Bang singularities, quantum fluctuations, DNA performance). It defined issues, ensuring to be particular and provides background examples on positions and arguments.
For example, when explaining theistic arguments for God’s existence, Qwen constructed a desk to make it simpler to grasp the premises, critiques, and proponents of probably the most related arguments.

Gemini matched this by protecting consciousness arguments that the majority fashions ignored. It additionally warned towards “God-of-the-gaps” reasoning extra explicitly than opponents.
ChatGPT introduced distinctive worth with its huge “Implications” part, exploring how the talk shapes science training coverage, bioethics legal guidelines, and private attitudes towards dying. This was much less tutorial and extra pragmatic, however nonetheless related to understand the character of the investigation.
Grok lined the foremost arguments however with much less element. It talked about fine-tuning and the anthropic precept, however did not cite particular values or talk about issues too deeply.
Outcome: Qwen and Gemini have been the most effective.
High quality: Taking all collectively—rigor, coherence, scholarly worth—which analysis would you wish to cite?
Each Qwen and Gemini produced studies you may undergo your professor. Qwen’s distinctive energy was balancing depth on each theistic traces and atheistic critiques, together with that burden-of-proof part. Gemini’s energy was integrating scientific frontiers (consciousness, evolution, cosmology) with philosophical arguments.
ChatGPT delivered substantial pedagogical worth—nice for educating or understanding implications. Grok labored as a dependable primer or fast reference.
In different phrases, ChatGPT and Grok are most likely those you’ll use in the event you simply wish to know one thing rapidly for a dialog, to impress your nerd date, or refresh your data earlier than a presentation on one thing you already know
Closing Scores:
- Qwen: 9/10
- Gemini: 9/10
- ChatGPT: 8/10
- Grok: 6/10
The podcast battle: Qwen vs Gemini
Qwen’s podcast function places it head-to-head with Google’s NotebookLM and Gemini, which pioneered AI-generated Audio Overviews.
In contrast to Gemini, Qwen affords a big number of host voices to select from. The construction is strong: two AI hosts have an precise dialog about your analysis, not only a text-to-speech read-through.

That mentioned, the voice high quality is inconsistent. Some voices are pure, however most of them sound robotic with bizarre accents. Throughout testing, one of many male hosts stored saying “oh oh oh” repeatedly, as a result of he was impressed. My spouse handed by and requested if I used to be watching porn.
With some trial and error, you’ll find a good voice that works easily, and the standard will increase significantly.
However Gemini and NotebookLM crush Qwen right here. Google’s Audio Overviews function—launched in NotebookLM in September 2024, expanded to Gemini in March 2025—sounds remarkably human. The speech patterns are pure, with back-and-forth banter and even humor.
Gemini’s podcasts really feel human and extra partaking.
Gemini additionally affords video generation, which is a big benefit for many who want an audiovisual strategy to understanding a subject quite than studying lengthy chunks of textual content.
Qwen can not do that—the truth is, no different mannequin can.
If you’d like full multimedia, together with audio, video, and internet, Gemini is probably the most full package deal.
The webpage benefit
Past analysis high quality, Qwen’s killer function is the auto-generated webpage. No different mannequin does this.
After your analysis finishes, you’ll be able to flip it right into a stay, hosted web site. Not a PDF or a Google Doc—an actual webpage with headers, formatted tables, embedded citations as hyperlinks.
The UI seems to be like Kimi; it options clear typography, responsive design and is immediately shareable.

ChatGPT customers have to repeat and paste into web site builders.
Gemini retains every little thing in Docs. Grok spits out textual content. Solely Qwen routinely generates web-ready output.
That workflow benefit is good to have.
Usually Clever Publication
A weekly AI journey narrated by Gen, a generative AI mannequin.
You might also like
More from Web3
Coinbase Files to List Single-Stock Perps on Apple, Tesla and Nvidia
Briefly Coinbase filed with the CFTC by means of Coinbase Derivatives to record single-stock perpetual futures within the US, searching …
Zcash Is Running—Devs Want to Make It Faster
In short Zcash builders are focusing on Nov. 5 to activate NU7, an improve that cuts block time—the interval between …
OpenAI Models Are Writing Their Own Jailbreak Instructions—And Sometimes Obeying Them
In short OpenAI revealed a brand new misalignment reporting framework alongside six experiences documenting regarding mannequin habits it discovered over …





