owndev
@owndev
·
11 days ago
·
a year ago
Google Gemini
Last Updated
11 days ago
Created
a year ago
Function
pipe
v1.16.0
Name
Google Gemini
Downloads
22K+
Saves
109+
Description
Highly optimized Google Gemini pipeline with advanced image and video generation capabilities, intelligent compression, and streamlined processing workflows.

Google Gemini Pipeline

Highly optimized Google Gemini pipeline with advanced image and video generation capabilities, intelligent compression, and streamlined processing workflows.


Features

  • Optimized asynchronous API calls for maximum performance
  • Intelligent model caching with configurable TTL
  • Streamlined dynamic model specification with automatic prefix handling
  • Smart streaming response handling with safety checks
  • Advanced multimodal input support (text and images)
  • Unified image generation and editing with Gemini 2.5 Flash Image Preview
  • Intelligent image optimization with size-aware compression algorithms
  • Automated image upload to Open WebUI with robust fallback support
  • Optimized text-to-image and image-to-image workflows
  • Non-streaming mode for image generation to prevent chunk overflow
  • Progressive status updates for optimal user experience
  • Consolidated error handling and comprehensive logging
  • Seamless Google Generative AI and Vertex AI integration
  • Advanced generation parameters (temperature, max tokens, etc.)
  • Configurable safety settings with environment variable support
  • Military-grade encrypted storage of sensitive API keys
  • Intelligent grounding with Google search integration
  • Vertex AI Search grounding for RAG
  • Native tool calling support with automatic signature management
  • URL context grounding for specified web pages
  • Unified image processing with consolidated helper methods
  • Optimized payload creation for image generation models
  • Configurable image processing parameters (size, quality, compression)
  • Flexible upload fallback options and optimization controls
  • Configurable thinking levels for Gemini 3 models with model-specific validation
  • Configurable thinking budgets (0-32768 tokens) for Gemini 2.5 models
  • Configurable image generation aspect ratio (1:1, 16:9, etc.) and resolution (1K, 2K, 4K)
  • Model whitelist for filtering available models
  • Additional model support for SDK-unsupported models
  • Video generation with Google Veo models (Veo 3.1, 3, 2)
  • Configurable video generation parameters (aspect ratio, resolution, duration)
  • Asynchronous video generation with progressive polling status updates
  • Automatic video upload to Open WebUI with embedded playback
  • Image-to-video generation support for Veo models
  • Negative prompt and person generation controls for video

Tools


GitHub

Open-WebUI-Functions

[!NOTE] For questions, problems, and feature requests, an issue must be created on GitHub.


License

Apache License 2.0


14