Google Gemini Pipeline
Highly optimized Google Gemini pipeline with advanced image and video generation capabilities, intelligent compression, and streamlined processing workflows.
Features
- Optimized asynchronous API calls for maximum performance
- Intelligent model caching with configurable TTL
- Streamlined dynamic model specification with automatic prefix handling
- Smart streaming response handling with safety checks
- Advanced multimodal input support (text and images)
- Unified image generation and editing with Gemini 2.5 Flash Image Preview
- Intelligent image optimization with size-aware compression algorithms
- Automated image upload to Open WebUI with robust fallback support
- Optimized text-to-image and image-to-image workflows
- Non-streaming mode for image generation to prevent chunk overflow
- Progressive status updates for optimal user experience
- Consolidated error handling and comprehensive logging
- Seamless Google Generative AI and Vertex AI integration
- Advanced generation parameters (temperature, max tokens, etc.)
- Configurable safety settings with environment variable support
- Military-grade encrypted storage of sensitive API keys
- Intelligent grounding with Google search integration
- Vertex AI Search grounding for RAG
- Native tool calling support with automatic signature management
- URL context grounding for specified web pages
- Unified image processing with consolidated helper methods
- Optimized payload creation for image generation models
- Configurable image processing parameters (size, quality, compression)
- Flexible upload fallback options and optimization controls
- Configurable thinking levels for Gemini 3 models with model-specific validation
- Configurable thinking budgets (0-32768 tokens) for Gemini 2.5 models
- Configurable image generation aspect ratio (1:1, 16:9, etc.) and resolution (1K, 2K, 4K)
- Model whitelist for filtering available models
- Additional model support for SDK-unsupported models
- Video generation with Google Veo models (Veo 3.1, 3, 2)
- Configurable video generation parameters (aspect ratio, resolution, duration)
- Asynchronous video generation with progressive polling status updates
- Automatic video upload to Open WebUI with embedded playback
- Image-to-video generation support for Veo models
- Negative prompt and person generation controls for video
Tools
GitHub
Open-WebUI-Functions
[!NOTE]
For questions, problems, and feature requests, an issue must be created on GitHub.
License
Apache License 2.0