baoyu-danger-gemini-web

Generate text and images via the Gemini Web API with vision input.

13|1|Updated Mar 4, 2026
One-click install
npx skills add https://github.com/ideacco/baoyu-skills-openclaw --skill baoyu-danger-gemini-web-ideacco
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill.
Skill: baoyu-danger-gemini-web
Source: https://github.com/ideacco/baoyu-skills-openclaw/tree/main/skills/baoyu-danger-gemini-web
Command: npx skills add https://github.com/ideacco/baoyu-skills-openclaw --skill baoyu-danger-gemini-web-ideacco

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes scripts (resource) and references (resource) components.

What problem does it solve?

This Skill provides a powerful AI backend for generating text and images, acting as a versatile tool for creative and informational content creation.

Core Features & Use Cases

  • Text Generation: Create written content from prompts, suitable for articles, summaries, or creative writing.
  • Image Generation: Produce images based on textual descriptions, ideal for illustrations, social media posts, or concept art.
  • Vision Input: Analyze and understand content from reference images to inform text or image generation.
  • Multi-turn Conversation: Engage in back-and-forth dialogue for iterative content refinement.
  • Use Case: Generate a blog post about sustainable living and then create accompanying featured images for the post.

Quick Start

Use the baoyu-danger-gemini-web skill to generate an image of a futuristic city at sunset.

Frequently Asked Questions about baoyu-danger-gemini-web

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I generate images from text prompts using Gemini Web API?▼

Multi-turn conversation allows iterative content refinement by engaging in back-and-forth dialogue. You can generate an initial text or image response and then provide follow-up prompts to update and improve the output.

Can I use reference images for vision input with AI generation?▼

Yes, you can use reference images for vision input. The Skill analyzes and understands content from provided images to inform subsequent text or image generation based on that visual context.

Do I need Node.js and Bun to run the Gemini Web Skill?▼

Yes, you need Node.js, Bun, and Chrome installed. These dependencies are strictly required for execution and authentication to access the reverse-engineered Gemini Web API functionalities.

What are the limitations of using a reverse-engineered Gemini Web API for image generation?▼

The main limitation of using a reverse-engineered Gemini Web API is potential instability, as it lacks official support and relies on Chrome for authentication, which may break with interface updates.

How does multi-turn conversation work for iterative content refinement?▼

Multi-turn conversation allows iterative content refinement by engaging in back-and-forth dialogue. You can generate an initial text or image response and then provide follow-up prompts to update and improve the output.