홈으로 돌아가기

Multi-Model AI Workspace — Chat, Text-to-Image & TTS

ForgeEcho docs for overseas ecommerce and marketing teams—prompt engineering, Amazon/Shopify listing photos with Nano Banana 2, UGC ad batch creatives, and TTS voiceovers in one Chat · Image · Voice pipeline.

Documentation Overview

ForgeEcho is a multi-model AI creative workspace for teams that ship assets every week: AI Chat for prompt engineering and script drafting, AI Image (Nano Banana 2 / Flux / GPT Image 2) for text-to-image and reference editing, and AI Voice for TTS narration. Chat → image → voice share one credit system so Amazon/Shopify listing work and Meta/TikTok creative tests stay on the same loop—not three disconnected apps.

The underlying loop is Plan → Evaluate → Improve: clarify intent in chat, score outputs against a short rubric, then promote winners into your Prompt Library. That discipline matters when you need consistent brand visuals, batch creative testing, or weekly social output—not a single lucky generation.

Why not ChatGPT + Midjourney + a TTS tab?

ApproachWhat breaks in production
Separate ChatGPT / Midjourney / ElevenLabs accountsPrompt history fragments; hard to hand off a “winner” to the next SKU
Discord-only image toolsWeak for reference-locked product photography and credit forecasting
One-shot prompting in any modelHigh re-roll cost; no shared evaluation rubric across the team

ForgeEcho keeps the same Plan–Evaluate–Improve discipline documented below, with credits and Prompt Library shared across Chat, Image, and Voice.

Who These Guides Are For

  • Solo creators publishing to TikTok, Reels, podcasts, and YouTube
  • Ecommerce teams refreshing listing photos and ad creatives at SKU scale
  • Marketing operators running voiceovers, poster campaigns, and social visuals
  • Anyone who wants a reliable chat → generate → iterate loop instead of one-off prompting

What You'll Learn

TopicOutcome
AI ChatTurn vague ideas into structured image prompts and voice scripts
AI ImageGenerate or edit visuals for products, ads, and creative assets
AI VoiceProduce ad reads, explainers, and social hooks from text
Style TransferShift to anime, cyberpunk, or cinematic looks while keeping composition
Ecommerce ImagesListing-ready photos with accurate color and clean backgrounds
Portrait RetouchNatural beauty enhancement without plastic skin
Batch SocialTemplate-driven weekly content with consistent brand look
Faceless ShortsChat → 9:16 stills → TTS → CapCut assembly
Credit & ModelsBudgets and when to use fast / 2K / 4K
Brand Visual SystemLocked prompt blocks + Prompt Library naming
Image TroubleshootingSymptom → one-lever fixes for common failures
Poster DesignCampaign layouts with clear typography zones
World Cup AI CreativeFan avatars, kicking poses, match commentary, memes
Optimize → GenerateThe core loop that connects Chat, Image, and Voice

Recommended Starting Path

  1. AI Chat Guide — clarify deliverables and draft prompts before generating
  2. Optimize Then Generate — adopt the five-step pipeline used across all scenarios
  3. Open the use-case guide that matches your next deliverable

AI Chat Guide

Multi-turn assistant for brainstorming, prompt drafting, and creative feedback.

How to Use AI Voice

Text-to-speech scripts and workflows for ads, explainers, and social audio.

How to Change Image Style with AI

Style transfer from reference image to finished asset.

Ecommerce Product Image Optimization

Listing photos, backgrounds, and conversion-ready variants with AI Image.

AI Photo Retouch and Beauty

Identity-preserving portrait retouch with AI Image.

Optimize Then Generate

The repeatable pipeline for Chat, Image, and Voice.

AI Poster Design Workflow

Posters and campaign covers with reusable prompts.

Social Media Batch Creative

Batch creatives for feed and short-form channels.

Faceless Short Video Pipeline

Scripts, 9:16 stills, TTS, and editor assembly for Shorts/Reels/TikTok.

Credit Budget & Model Selection

Unit costs, weekly budgets, and when to use fast vs 2K vs 4K.

Brand Visual System

Prompt Library naming and locked blocks for consistent brand look.

AI Image Troubleshooting

Fix shape drift, garbled labels, plastic skin, and bad crops fast.

World Cup AI Creative

Fan avatars in national kits, kicking poses, AI match commentary, posters, and memes.

How the Guides Connect

Every scenario guide assumes the same underlying workflow: draft in AI Chat (0.5 credits per reply), refine until the prompt or script is specific enough, generate in AI Image (3–8 credits by resolution) or AI Voice (1 credit per 500 characters), then iterate on one variable at a time. Optimized prompts are saved automatically to your dashboard Prompt Library for reuse across batch runs.

If you are new to generative tools, start with Optimize Then Generate—it explains why vague prompts fail and how structured instructions improve stability across Nano Banana 2, Flux-family models, and TTS voices.

Workspace Quick Reference

ModuleTypical useCredits (approx.)
AI ChatPrompt engineering, UGC scripts, failure diagnosis0.5 per reply
AI ImageText-to-image, reference edit, 4K listing assets3–8 per image by resolution
AI VoiceAd reads, short-form hooks, explainers1 per 500 characters

The 2026 multi-model creative stack is intentional: Chat for prompt engineering, Nano Banana 2 for text-to-image, and TTS for narration. Chat models (Gemini family) and image models (Nano Banana 2 family) play different roles—chat interprets and rewrites; image executes pixels. Swapping those roles usually costs more credits than following the pipeline.

30-minute quick start

If you only have half an hour, run this minimal loop once:

  1. Open AI Chat guide and generate one structured image prompt from the sample request (~5 min)
  2. Paste into AI Image, pick nano-banana-fast + 1K, generate 2 variants (~10 min)
  3. Score with the checklist in Optimize Then Generate, change one variable, regenerate once (~10 min)
  4. Save the winning prompt to your library with aspect ratio and model notes (~5 min)

After one pass, open the scenario card above that matches your next deliverable.

Pick the right guide (by deliverable)

Your next taskStart hereWhy
Listing photos, white background, color accuracyEcommerce Image OptimizationPlatform specs, A/B testing, batch SKU workflow
LinkedIn headshot, dating/social portraitPhoto Retouch and BeautyIdentity-preserving two-pass retouch
Anime, cyberpunk, cinematic restyleImage Style TransferReference upload + preservation blocks
Event promo, sale banner, campaign coverAI Poster Design WorkflowTypography safe zones and ratio exports
Weekly feed/stories at volumeSocial Media Batch CreativeTemplates, credit planning, fatigue testing
Faceless Shorts / Reels / TikTokFaceless Short Video PipelineChat → stills → TTS → CapCut
Credits melting / wrong model choiceCredit Budget & Model SelectionBudgets and stage-based models
Brand look drifts across operatorsBrand Visual SystemLocked blocks + library naming
Bad outputs, unsure what to changeAI Image TroubleshootingSymptom → one-lever fixes
Ad read, explainer, social hook audioAI Voice WorkflowScript structure, pacing, voice selection
Prompts feel vague or outputs driftOptimize Then GeneratePlan–Evaluate–Improve loop before any scenario

Model selection at a glance

GoalChat modelImage modelResolution
Quick prompt draftAny Gemini / chat model——
Fast visual filter (6+ variants)—nano-banana-fast1K
Final listing / ad still—nano-banana-22K
Print or large display—nano-banana-2 or nano-banana-pro4K
Voice tone test——Short clip first

Chat interprets intent; image models render pixels. Run at least one chat round before spending image credits on a new scenario.

When outputs fail (quick diagnosis)

SymptomLikely causeFix (one lever)
Wrong colors / plastic skinOver-aggressive prompt or no referenceUpload reference; reduce beauty words; split lighting vs skin passes
Composition breaks at 9:16Prompt composed for landscapeRewrite prompt for vertical safe zones
Face doesn't match photoText-only generationUpload reference; add "preserve identity, same person"
Garbled text on productModel inventing labelsSimplify label expectations; overlay text in design tool
Robotic voice readScript written for reading, not speakingRe-draft in AI Chat with "spoken phrasing, short sentences"
Inconsistent batch lookPrompt drift between runsSave winner to Prompt Library; change one variable only

For the full evaluation rubric, see Optimize Then Generate.

Documentation OverviewWhy not ChatGPT + Midjourney + a TTS tab?Who These Guides Are ForWhat You'll LearnRecommended Starting PathHow the Guides ConnectWorkspace Quick Reference30-minute quick startPick the right guide (by deliverable)Model selection at a glanceWhen outputs fail (quick diagnosis)