Skip to content

Features Overview ​

Go Bananas! provides a comprehensive set of features for AI-powered image generation and management.

Feature Categories ​

Models & Best Practices ​

FeatureDescription
Pro Model Best PracticesGoogle's 10 tips for professional generation
OpenAI gpt-image-2OpenAI's flagship image model — text rendering, prompt adherence, quality tiers
OpenAI GPT Image 2.5Flare and Sunburst — xhigh/max quality and transparent backgrounds

Golden Rules of Prompting

  1. Edit, Don't Re-roll - Use continue_editing for 80%+ correct images
  2. Natural Language - Full sentences, not tag soups
  3. Materiality - Describe textures: "brushed steel", "soft velvet"
  4. Context - Add "for whom": changes everything
  5. Identity Locking - "Keep facial features exactly the same as Image 1"

Image Creation & Editing ​

FeatureDescription
Image GenerationCreate images from text with style control
Image EditingConversational editing without re-uploading

Library Management ​

FeatureDescription
CharactersConsistent characters across scenes
ProductsMarketing images with product consistency
Workflow DiagramsMermaid → professional visual diagrams & infographics
Style PresetsReusable style templates

Session & Analytics ​

FeatureDescription
SessionsConversation context for workflows
AnalyticsUsage tracking and quota management

Feature Comparison ​

Provider comparison hero

Six models from two providers, one API. Pick per request via `model_id`.

Provider Comparison ​

CapabilityLite (Nano Banana 2 Lite)Flash (Nano Banana 2)Pro (Go Bananas! Pro)OpenAI gpt-image-2OpenAI GPT Image 2.5 (Flare, Sunburst)
Default✅————
ProviderGemini 3.1 Flash LiteGemini 3.1 FlashGemini 3 ProOpenAIOpenAI
Generation speedFastestFastQualityQualityFlare: fast; Sunburst: precise
Max prompt length16,384 chars16,384 chars16,384 chars32,000 chars32,000 chars
Max reference images1414141616
Resolution options1K0.5K / 1K / 2K / 4K tiers1K / 2K / 4K tiersAny size meeting constraints (≤3,840px edge, multiples of 16, ≤3:1)Same as GPT Image 2
Aspect ratios13 standard/cinematic14 API ratios, including 1:4–8:1 extremes10 API ratios, including 21:9 and social formats13 (incl. 21:9 cinematic, 4:5/5:4 social)13 (same as GPT Image 2)
Output formatsPNGPNGPNGPNG, JPEG, WebP + output_compressionPNG, JPEG, WebP + output_compression
Quality levelsautoautoautolow / medium / high / autolow / medium / high / xhigh / max / auto
Google Search groundingNoYesYesNoNo
Thinking modeYesYesYesNot exposedNot exposed
Transparent backgroundNoNoNoNot supported by modelYes (PNG or WebP)
Text renderingBasicGoodAdvanced (multilingual)Best-in-class — small text, dense layoutsBest-in-class
Identity lockingLimitedStrongFull (up to 5 people)Strong via reference editsStrong via reference edits
Default request timeout30s60s120s240s240s
CostLowestLowMidToken-priced by input and outputToken-priced by input and output

Pro Model Exclusive Features ​

FeatureDescription
Text & InfographicsSOTA text rendering for posters, menus, data visualizations
Viral ThumbnailsIdentity + text + graphics in one pass
StoryboardingMulti-scene sequential art with consistent characters
Structural ControlSketch-to-final, wireframe-to-UI workflows
2D↔3D TranslationFloor plans to renders, 2D art to 3D
High-Fidelity UpscalingInput images as small as 150x150 enhanced to 4K
Photo RestorationRepair tears, scratches, fading in old photographs
360 TurnaroundGenerate multiple angles from a single reference image
Thinking ModeEnhanced reasoning for complex prompt interpretation
Google Search GroundingReal-time data and factual accuracy via web search
Person Generation SafetyConfigurable safety modes: dont_allow, allow_adult, allow_all

See Pro Model Best Practices for complete guide.

When to Use Each Model ​

Use Lite (default) for:

  • Quick iterations and drafts
  • Simple prompts
  • High-volume generation
  • Cost-sensitive workflows

Use Flash for:

  • Broader aspect ratios
  • 2K/4K output
  • Stronger multi-reference workflows
  • Search-grounded generation

Use Pro for:

  • Final production assets
  • Print materials (4K resolution)
  • Text-heavy content (infographics, posters)
  • Multi-person character consistency
  • Complex multi-image compositions

Feature Matrix ​

FeatureMCP ToolREST APIWeb Console
Generate images✅✅✅
Edit images✅✅✅
Continue editing✅✅✅
Characters✅✅✅
Multi-character✅✅✅
Products✅✅✅
Style presets✅✅✅
Search images✅✅✅
Usage stats✅✅✅
Bulk operations❌✅✅
Download ZIP❌✅✅
Admin functions❌✅✅

Quick Examples ​

Generate an Image ​

Generate a cozy cabin in snowy mountains at sunset

Edit Conversationally ​

Add smoke coming from the chimney
Make the sky more dramatic
Add some deer in the foreground

Create and Use a Character ​

Create a character "Forest Guardian" - an ancient tree spirit
with glowing green eyes and bark-like skin covered in moss

Generate Forest Guardian protecting a hidden waterfall
Generate Forest Guardian teaching woodland creatures

Use a Style Preset ​

Create style preset "Anime" with prompt prefix:
"anime style, vibrant colors, detailed shading"

Generate a warrior princess using the Anime style

Integration Patterns ​

Best for conversational AI workflows:

  • Claude.ai / Claude Desktop remote connector
  • Claude Code
  • Codex
  • Claude Desktop
  • Cursor IDE
  • Custom MCP clients

Use remote HTTP OAuth when your client supports it. Use the gobananas-mcp STDIO proxy for command-based or headless clients.

REST API ​

Best for programmatic access:

  • Custom applications
  • Batch processing
  • Webhooks and integrations

Web Console ​

Best for visual management:

  • Browse and search images
  • Manage characters and presets
  • View analytics
  • Admin operations

Next Steps ​

Dive into each feature:

  1. Pro Model Best Practices → (Start here!)
  2. Image Generation →
  3. Image Editing →
  4. Characters →
  5. Products →
  6. Workflow Diagrams →
  7. Style Presets →
  8. Sessions →
  9. Analytics →

Released under the MIT License.