Qwen logo

Qwen output UX: images, video, search & selection

Updated August 25, 2026

Qwen threads treat media and search as continuations, not side quests. Upload a portrait and the model writes a long image prompt before generate. Finished images get Create Video and Edit overlays. Say "animate it" and the bar picks up a Create Video chip with the source thumbnail. Web search answers ship with Searched the web, inline domains, a Thinking and Search sidebar, and selection actions (Copy, Ask Qwen, Explain, Translate). The UI looks trustworthy. In this capture it also confidently cites a fictional Apple MacBook Neo for 2026, which is the trust gap the chrome hides.

Photo to image prompt

Upload a portrait, arm Create Image mode, and send without typing a prompt.
Upload a portrait, arm Create Image mode, and send without typing a prompt.

What works

  • The assistant reply is a full art-direction brief (yarn doll, lighting, texture) instead of jumping straight to pixels.
  • User sees the plan in chat before spend. Easier to edit text than regenerate blindly.
  • Create Image chip and Qwen-Image 2.0 / 16:9 stay pinned in the composer through the reply.

What we would push on

  • The prompt is long and single-block. No highlight of what changed from the upload.
  • Send stays greyed while mode chips are set. Unclear if user must approve the text prompt first.

Product bet

Showing prompt expansion builds trust for image spend. Users feel the model understood the reference photo before GPU time runs.

Tradeoff

DecisionBenefitCost
Visible text prompt step before image generateEdit scope before render; teaches prompt qualityExtra scroll; slower time-to-image

Takeaway

For reference-image gen, show the rewritten prompt in-thread and let users edit it before render.

Image in thread

After generate, view the image in the thread with action overlays.
After generate, view the image in the thread with action overlays.

What works

  • Image renders inline at readable size with share, regenerate, and overflow on the message.
  • Create Video and Edit float on the image. Next steps are on the artifact, not buried in +.
  • Thumbs, share, and download also sit top-right on hover for quick feedback and export.

What we would push on

  • Composer still shows 16:9 while the output image is square. Format chip and result can disagree.
  • AI-generated content may not be accurate disclaimer is easy to miss under the bar.

Tradeoff

DecisionBenefitCost
Inline image with Create Video / Edit overlays plus message actionsChains image to video without re-uploadAspect chip can lie; disclaimer fades into footer

Takeaway

Put chain actions on the image itself. Match aspect label to what actually rendered.

Animate with reference

With the image in thread, type "animate it" in the composer.
With the image in thread, type "animate it" in the composer.

What works

  • Source thumbnail appears in the bar above animate it. Context is visible before send.
  • Create Video chip replaces Create Image automatically. Mode follows intent.
  • Short verb prompt works. User did not re-describe the doll.

What we would push on

  • No motion controls (duration, camera move) in the bar yet, only 16:9 later.
  • Chip swap is smart but silent. A one-line "Switching to video from this image" would teach the pattern.

Product bet

Image-to-video is the upsell path. Reference thumbnail plus auto chip swap keeps users in one thread instead of opening a video app.

Tradeoff

DecisionBenefitCost
Reference thumbnail in bar with auto Create Video chipNatural language handoff; no re-uploadLimited motion controls at prompt time

Takeaway

When a reply references prior media, pin the thumbnail and swap mode chips to match the next job.

Video generation progress

Send the animate request and wait for the progress card.
Send the animate request and wait for the progress card.

What works

  • Gradient card with spark icon and 4% label. Users know video is rendering, not stuck.
  • Create Video chip and 16:9 stay in the composer so settings survive the wait.
  • User bubble animate it stays minimal. Focus is on the job card.

What we would push on

  • No ETA or stage name beyond percent. Long video jobs will feel opaque.
  • Send button greys out. Cannot queue a follow-up until render finishes.

Takeaway

Show percent for video gen, but add stage copy or ETA when renders run longer than a few seconds.

Preview, download, publish

Open the finished video preview modal.
Open the finished video preview modal.

What works

  • Modal player with 00:00 / 00:05 timer. Short clip length is obvious before download.
  • Download and Publish are equal primary actions under the player.
  • Create Video chip and 16:9 remain in the bottom bar for another pass without closing context.

What we would push on

  • Publish destination is not named on the button. Trust requires knowing where it goes.
  • Square video in a 16:9 modal adds letterboxing. Same aspect mismatch as image output.

Tradeoff

DecisionBenefitCost
Modal preview with Download and Publish plus persistent mode barClear export path; easy retryPublish target vague; aspect mismatch visible

Takeaway

Treat video like a deliverable: preview, duration, download, and name the publish target on the button.

Web search answer

Send Best budget laptops with Web search mode armed.
Send Best budget laptops with Web search mode armed.

What works

  • Lead sentence states intent: I'll search for information about the best budget laptops…
  • Searched the web row is tappable. Signals retrieval happened.
  • Inline grey domain chips after claims (pcworld.com, pcmag.com, cnet.com) mirror Perplexity-style trust chrome.
  • Structured list with bold product names and price line reads scannable.

What we would push on

  • Answer cites 2026 reviews and an Apple MacBook Neo at $599. That product does not exist. The UI still looks authoritative.
  • Domains are not numbered footnotes. Hard to match a claim to a specific source click.
  • Disclaimer under the bar is tiny compared to the confident list above it.

Product bet

Search mode competes on Perplexity-shaped answers. Citations and Searched the web sell freshness even when the model fills gaps with fiction.

Tradeoff

DecisionBenefitCost
Chat-formatted search answer with inline domain chipsFamiliar reading flow; looks researchedHallucination wears a citation costume; weak claim-to-source link

Takeaway

If you show domain chips, tie each claim to a numbered source and surface when retrieval failed or dates look wrong.

Thinking and Search panel

Open the Thinking and Search sidebar on a web search answer.
Open the Thinking and Search sidebar on a web search answer.

What works

  • Sidebar title Thinking and Search sets expectation: reasoning plus retrieval, not chat history.
  • Source pills (wired.com, cnet.com, pcworld.com) and +14 show breadth at a glance.
  • Numbered results include page titles and snippets so users can sanity-check before trusting the summary.

What we would push on

  • Snippets also reference 2026 and MacBook Neo. Transparency reveals the same hallucination problem, not a fix.
  • Panel is optional. Many users will never open it and only see the polished main answer.

Tradeoff

DecisionBenefitCost
Optional sidebar with source list and overflow countPower users can audit retrievalAudit surface still bad if sources are wrong; easy to ignore

Takeaway

A source drawer helps only if snippets are real. Pair it with date checks and a low-confidence state in the main answer.

Selection actions

Highlight text in the search answer (e.g. Price: $599).
Highlight text in the search answer (e.g. Price: $599).

What works

  • Floating menu: Copy, Ask Qwen, Explain, Translate with a language submenu.
  • Translate lists many locales in one scroll (English US, Arabic, German, Spanish, Persian, French, Hindi, Italian, Japanese, Korean…).
  • Selection works on search answers, not only plain chat prose.

What we would push on

  • Ask Qwen vs Explain overlap. Users will not know which to pick.
  • No Check sources or Cite selection action despite search mode being on.

Tradeoff

DecisionBenefitCost
Generic selection menu with translate submenu on search outputReuse chat refinement patterns on research answersMissing search-specific actions; Ask vs Explain redundant

Takeaway

Extend selection menus for search threads with source-check actions, not only translate and explain.

Follow-ups and source badges

Scroll to the bottom of the search answer for chips and source globes.
Scroll to the bottom of the search answer for chips and source globes.

What works

  • Three follow-up pills continue the research thread (under $500, battery life, what to look for).
  • Message actions row includes copy, feedback, share, regenerate, and globe badges with +17.
  • Web search chip persists in the composer so the next turn stays in search mode.

What we would push on

  • Follow-ups repeat the 2026 framing. Bad suggestions reinforce bad facts.
  • +17 globes are not clickable domains in this view. Count impresses more than it informs.

Tradeoff

DecisionBenefitCost
Follow-up pills plus source count badges on search answersKeeps research session going; signals many sourcesSuggestions can amplify errors; badge count without list

Takeaway

Follow-ups should deepen verification, not repeat the same shaky premise. Make source badges open the drawer.

Copy this

  • Prompt expansion in-thread before image generate from a reference photo
  • Create Video and Edit overlays on finished images
  • Reference thumbnail plus auto Create Video chip on animate it
  • Video progress card with percent while mode chips stay in the composer
  • Preview modal with Download and Publish for short clips
  • Searched the web row plus inline domain chips on search answers
  • Thinking and Search sidebar with numbered snippets and +N overflow
  • Selection menu with Translate submenu on research output
  • Follow-up pills that match an armed Web search composer chip

Skip this

  • Citation chips on answers that cite fictional products
  • 16:9 selected while square image or letterboxed video renders
  • Publish button with no destination named
  • Ask Qwen and Explain both on selection without guidance
  • Follow-ups that double down on wrong dates or products
  • +17 source globes that do not open a source list

How others output, artifacts & refinement

How other products handle the same job, and what each tradeoff reveals.

Compare output & artifacts UX

ChatGPT, Claude, Perplexity, and Gemini side by side: refinement, artifacts, export, and patterns to copy.

Full comparison

Useful for a critique or spec? Share it.

Original gallery pages: Output & Refinement