Agent Skills: gemini-3-multimodal

Process multimodal inputs (images, video, audio, PDFs) with Gemini 3 Pro. Covers image understanding, video analysis, audio processing, document extraction, media resolution control, OCR, and token optimization. Use when analyzing images, processing video, transcribing audio, extracting PDF content, or working with multimodal data.

UncategorizedID: adaptationio/skrillz/gemini-3-multimodal

Install this agent skill to your local

pnpm dlx add-skill https://github.com/adaptationio/skrillz/gemini-3-multimodal

Skill Files

Browse the full folder contents for gemini-3-multimodal.

Download Skill

Loading file tree…

Select a file to preview its contents.