Gemma 4 models leaking "think" capability #15267
Unanswered
61150n4n1m35-svg
asked this question in
Q&A
Replies: 1 comment 1 reply
|
You used a different prompt in the two examples. |
1 reply
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment



Uh oh!
There was an error while loading. Please reload this page.
I would like to report a possible bug or issue regarding how ComfyUI loads the Gemma 4 model. Recent updates have changed the way Gemma 4 models are loaded; up to version 0.25, the model worked perfectly as an LLM for vision tasks—such as describing images or refining/expanding prompts—outputting only the description or the expanded prompt. However, starting with version 0.29, the behavior changed: the model now "leaks" its thought process, acting more like a general assistant and ignoring system prompt instructions by including its internal reasoning in the output, even when the system prompt explicitly forbids this. In the images below, I show the two versions tested in Stability Matrix, along with screenshots of the results: the correct description from version 0.25 versus the description containing
<think>tags from version 0.30. Both tests used the same workflow (from ComfyUI's built-in templates), the recommended Gemma 4 model, and identical prompts and images, yet yielded vastly different results.Tested versions

Result from ComfyUI current version 0.30

Result from ComfyUI version 0.25

Here I am comparing the two results generated in the same workflow across two different versions of ComfyUI.
Result in ComfyUI version 0.25
This is a vibrant and energetic piece of fan art featuring the character Hitori "Bocchi" Gotoh from the anime and manga series Bocchi the Rock!.
Here is a detailed description:
Character:
Composition and Background:
Mood and Style:
Result in ComfyUI version 0.30 (current)
Here's a thinking process to arrive at the detailed image description:
Analyze the Request: The user wants a description of the provided image.
First Impression & Subject Identification:
Detailed Breakdown - The Character:
Detailed Breakdown - The Background & Text:
Synthesis and Structuring the Description: A good description moves from general to specific.
Refining the Language (Self-Correction/Enhancement):
Final Polish: (This leads directly to the polished response provided below.)This is a vibrant, dynamic illustration in the style of modern anime, heavily themed around rock music. The central figure is a character strongly resembling Hitori "Bocchi" Gotoh from the series Bocchi the Rock!.
The Character:
The character has long, voluminous, bright pink hair and large, expressive blue eyes. She is depicted in a highly energetic pose, mid-performance. She is dressed in a casual but edgy outfit: a dark (likely black) t-shirt featuring a graphic design, paired with a short, flowing skirt that incorporates shades of dark blue and purple. She wears dark socks and sturdy, dark shoes laced with red string.
She is actively playing or posing with a sleek, black electric guitar, which is held across her body. Her posture is confident and lively, conveying the energy of a live musician.
The Background and Atmosphere:
The background is dominated by deep, earthy brown and sepia tones, giving the scene a gritty, raw, and atmospheric feel, reminiscent of a dimly lit concert venue or a worn stage.
Overlaying the top portion of the image is bold, distressed typography in bright red. Large, impactful lettering spells out "BOCCHI!" and fragments of other text, such as "...ROCK!", clearly branding the artwork to the musical theme of the character.
Overall Impression:
The image successfully captures the blend of shy vulnerability (implied by the character's usual demeanor) and explosive musical passion, using strong color contrast (bright pink and blue against the dark browns and red text) to create a high-energy, rock-and-roll aesthetic.
All reactions