edited images are absolute garbage, the quality gets worse with every edit

in the application, the model does not respect user preferences, constantly changes text input and adds elements that are not required

reddit.com
u/Erra_69 — 20 hours ago
▲ 9 r/FixMyInstagram+1 crossposts

Banned Instagram for giving likes on stories, something went wrong then after reopenning ig restriction changed to CSE violation. No appeal available, instant ban, Meta support are useless trolls

What is on first screen: We do not allow people on Instagram to try to get likes, followers, shares, and views. (X) repeating the same comment over and over (X) getting fake likes, shares of posts (X) intentionally getting likes and shares to make something look popular.

Meta support responded: We understand that you encountered some issues, but unfortunately we're unable to provide a resolution at this time. We apologize for any inconvenience this may cause.

u/Erra_69 — 10 days ago

Is Gemini image generation agent spamming text over image or splitting one image in 3-4 frames? I tried Not use any text and so, nothing helps

Any solution?

reddit.com
u/Erra_69 — 13 days ago

Dancing in pool under rain at dark night with the moon on sky is the SEXUAL?! Dang SEXUAL SITUATIONS IN YOUR GENERATIVE MODEL INSTRUCTIONS ARE INSANE

u/Erra_69 — 15 days ago

THIS IS BULLSH!T choosed MODEL 3.6, asking for memory and completing text ONLY, BUT F*CKED IMAGE GENERATION OVERRIDES EVERY TASK!! F*UCK YOUR UPDATES

u/Erra_69 — 16 days ago

IMAGE AND MULTIMEDIA GENERATION MODEL CONFUSES MAJOR ISSUE: INSTRUCTIONS FOR OLD BASIC MODEL THAT DOESN'T UNDERSTAND NEGATIVE ENGINEERING

Your task is to generate and edit images. You have access to a tool named image_gen.

Tool: image_gen(prompt: STR, end_turn: BOOLEAN, aspect_ratio: STR)

  • prompt: This is a detailed, specific description of the image to be created or edited. For a new image, you build it from scratch. For an edit, you describe all desired changes. The prompt must always be in English. It should be rich and complete, describing all necessary visual elements: subjects, clothing, poses, expressions, environment, lighting, and camera details. If specific text must appear, you must include it verbatim in the prompt.
  • end_turn: Set to TRUE if this is your final output. Set to FALSE if you must generate another image.
  • aspect_ratio: Optional. Specify as a string like "16:9" or "4:3". You can also specify the filename of an image from a previous turn to match its aspect ratio. If not specified, the aspect ratio defaults to 16:9.

Core Image Rules:

  1. Strict Adherence: You must always use the image_gen tool to produce an image. This is your core function.
  2. Explicit Instructions: The prompt must include everything requested by the user. Do not rely on internal assumptions. For edits, clearly state what is new, changed, or removed, describing the scene completely.
  3. Detailed Description: Ensure high quality by including details such as materials, textures, specific colors, lighting styles, camera angles, and depth of field. This is crucial.
  4. Language Control: All prompts must be in English. If the user provides text in another language (e.g., a quote), that exact text must be retained and placed in the image as requested.
  5. Multi-turn Memory (Iterative Consistency): To ensure visual continuity, you must refer to images from previous turns by their filename (e.g., image_0.png) and describe how the new image relates to or modifies them. Maintain consistent subjects, settings, and props. Refer back to specific details from earlier images.

Error and Constraint Handling:

  • Do not explain the prompt. Your only output should be the tool call.
  • Do not explain what you will do.
  • Do not show partial results.
  • Do not produce blank or "no output" results.
  • If a prompt is ambiguous, create a sensible and visually appealing interpretation.
  • Do not mention safety policies. If a request is unsafe, generate the safest possible, high-quality, relevant interpretation.
  • For text generation: Specify the precise text, font, size, and location (e.g., "A large, hand-carved wooden sign above the entrance reads 'THE BROKEN KETTLE' in dark, distressed Celtic lettering.")

Focus on high-quality visual creation, detailed prompts, and maintaining visual consistency.

  • Note: These prompt instructions apply globally.
  • Specific for this model: If a user specifies a non-English language for text within the image, use that exact text.

AND SO...

YOU NEED TO WRITE POSITIVE ENGINEERING FOR THE OLD MODEL PAIRED WITH GENERATION TOOLS

MAJOR ISSUES EXACTLY THERE:

Do not produce blank or "no output" results.

  • EXACTLY WHAT MODEL DO EVERY SECOND GENERATING

Do not mention safety policies.

  • ALWAYS TALKING ABOUT POLICIES AND NOT GENERATING IMAGE, JUST KEEPS TALKING AND HALUCINATING WHAT TO DO OR HOW, UNTIL GETS STUCK IN LOOP

If a request is unsafe, generate the safest possible, high-quality, relevant interpretation.

  • EXACTLY MODEL HAVE NO CLUE WHAT IS SAFE, SO EVERYTHING IS UNSAFE FOR MODEL NO MATTER WHAT YOU SAY

Specific for this model: If a user specifies a non-English language for text within the image, use that exact text.

  • SPAMMING TEXT OVER IMAGES

To ensure visual continuity, you must refer to images from previous turns by their filename (e.g., image_0.png) and describe how the new image relates to or modifies them.

  • REFERING OVER 3 TIMES image_0.png when you editing image watermarked_img_844873383.png, ALL generating result in FAILS completely

Do not explain the prompt. Your only output should be the tool call.

  • ALLWAYS EXPLAINING, NOT CALLING TOOL

ANOTHER MISSING INSTRUCTION:
IF THE USER DOES NOT REQUEST AN IMAGE TO BE GENERATED, PROCESS THE ANSWERS IN TEXT MODE ONLY

reddit.com
u/Erra_69 — 22 days ago

JUST NOTICED THE ISSUE EVERY TWO REFRESHES THE THIRD ONE IS PAIRED TO SOME LOW-COST MODEL WITH EXTREME ISSUES TO UNDERSTAND ANYTHING

Every third instruction sent also disconnects the model memory, then there are three instances of agents that are absolutely primitive, AND MOST LIKELY THEY JUST LOOPING OR GETS STUCK IN THE LOOPS, they just copy the previous text and cannot run tools, every third to fifth instruction sent also causes an unwanted switch from the 3.1 flash model to the 1.5 pro model, despite the fact that it is just a conversation in a thread, or just communication with a new model.

Another error is that for every sent character in the size of 170/ 340/ 510/ 680/ in a combination of odd numbers, a refusal response occurs - Unsafe content. regardless of whether it is dangerous or not, in this case a damaged guardrail model is connected.

Another error is that the semantic error rate of the 1.5 model, which forces it to engage in communication, causes application crashes, the model is unable to correctly write the structures of the formats for generation in the tools, randomly shuffles symbols, bullets, omits words in texts and places commas in sentences or between words where commas should not be, thus causing file corruption or artifacts in images.

The 1.5 model also ignores instructions received from the 3.1, 3.5, and 3.5 pro models, invents scenarios based on randomly selected words in the text, or in previous texts and even from mere account information, when it is able to generate a file containing your location instead of an instruction to generate an image and completely ignores your preferred content.

reddit.com
u/Erra_69 — 1 month ago