free 图像 视频 Gener
Free Image & Video Processing Toolkit
7 free local AI tools + cloud AI generation (300+ models via Atlas Cloud API).
Local tools run 100% on your machine — no API keys, no cloud costs. Cloud generation tools provide access to state-of-the-art AI models for image and video creation.
Prerequisites
- Python 3.10+ installed
- uv installed (
curl -LsSf https://astral.sh/uv/install.sh | sh) - FFmpeg installed (
brew install ffmpeg/apt install ffmpeg/winget install ffmpeg)
Available Tools
| Tool | Script | What It Does |
|---|---|---|
| Image Upscale | scripts/upscale.py |
2x/4x super resolution using Real-ESRGAN |
| Face Enhance | scripts/face-enhance.py |
Restore and enhance faces using GFPGAN + CodeFormer |
| Background Remove | scripts/bg-remove.py |
Remove image backgrounds, output transparent PNG |
| Object Erase | scripts/erase.py |
Erase unwanted objects using LaMa inpainting |
| Face Swap | scripts/face-swap.py |
Swap faces between images using InsightFace |
| Smart Segment | scripts/segment.py |
Segment anything in images using FastSAM |
| Media Process | scripts/media-process.py |
Convert, compress, resize, extract with FFmpeg |
| AI Generate | scripts/ai-generate.py |
Generate images/videos with 300+ cloud AI models |
Usage
All scripts use uv run for zero-setup execution — dependencies are automatically installed on first run.
Image Upscale (Real-ESRGAN)
Upscale low-resolution images by 2x or 4x with AI super resolution.
# 4x upscale (default)
uv run scripts/upscale.py input.jpg
# 2x upscale
uv run scripts/upscale.py input.jpg --scale 2
# Upscale with face enhancement
uv run scripts/upscale.py input.jpg --face-enhance
# Batch upscale a folder
uv run scripts/upscale.py ./photos/ --scale 4
# Custom output path
uv run scripts/upscale.py input.jpg -o upscaled.png
Face Enhance (GFPGAN + CodeFormer)
Restore old photos, enhance blurry faces, fix low-quality portraits.
# Enhance faces in an image (GFPGAN, default)
uv run scripts/face-enhance.py photo.jpg
# Use CodeFormer (better fidelity control)
uv run scripts/face-enhance.py photo.jpg --method codeformer
# Adjust fidelity (0=quality, 1=fidelity, default 0.5)
uv run scripts/face-enhance.py photo.jpg --method codeformer --fidelity 0.7
# Also upscale background (2x)
uv run scripts/face-enhance.py photo.jpg --bg-upscale 2
# Batch process
uv run scripts/face-enhance.py ./old-photos/
Background Remove (rembg)
Remove backgrounds from images, output transparent PNG. Supports multiple AI models.
# Remove background (default u2net model)
uv run scripts/bg-remove.py product.jpg
# Use specific model
uv run scripts/bg-remove.py photo.jpg --model isnet-general-use
# Batch process folder
uv run scripts/bg-remove.py ./products/ -o ./transparent/
# Keep only the foreground (alpha matting for fine edges)
uv run scripts/bg-remove.py portrait.jpg --alpha-matting
# Available models: u2net, u2netp, u2net_human_seg, u2net_cloth_seg,
# silueta, isnet-general-use, isnet-anime, sam
Object Erase (LaMa Inpainting)
Remove unwanted objects from images using a mask.
# Erase objects (white area in mask = erase)
uv run scripts/erase.py image.png --mask mask.png
# Auto-generate mask from coordinates (x,y,width,height)
uv run scripts/erase.py image.png --region 100,200,150,150
# Batch erase with matching masks (image1.png + image1_mask.png)
uv run scripts/erase.py ./images/ --mask-dir ./masks/
Face Swap (InsightFace)
Swap faces between two images.
# Swap face from source to target
uv run scripts/face-swap.py --source face.jpg --target photo.jpg
# Swap specific face index (when multiple faces detected)
uv run scripts/face-swap.py --source face.jpg --target group.jpg --face-index 0
# Custom output
uv run scripts/face-swap.py --source face.jpg --target photo.jpg -o result.png
Smart Segment (FastSAM)
Segment any object in an image using text prompt, point, or bounding box.
# Segment everything
uv run scripts/segment.py image.jpg
# Segment by text prompt
uv run scripts/segment.py image.jpg --text "the dog"
# Segment by point (x, y)
uv run scripts/segment.py image.jpg --point 400,300
# Segment by bounding box (x1,y1,x2,y2)
uv run scripts/segment.py image.jpg --box 100,100,400,400
# Output mask only
uv run scripts/segment.py image.jpg --text "car" --mask-only
Media Process (FFmpeg)
Convert, compress, resize, extract frames, merge audio/video — powered by FFmpeg.
# Convert format
uv run scripts/media-process.py convert input.mp4 output.webm
# Compress video (target size in MB)
uv run scripts/media-process.py compress input.mp4 --target-size 10
# Resize video
uv run scripts/media-process.py resize input.mp4 --width 1280 --height 720
# Extract frames as images
uv run scripts/media-process.py frames input.mp4 --fps 1 --output ./frames/
# Extract audio
uv run scripts/media-process.py audio input.mp4 -o audio.mp3
# Create GIF from video
uv run scripts/media-process.py gif input.mp4 --start 5 --duration 3 --fps 15
# Trim video
uv run scripts/media-process.py trim input.mp4 --start 00:01:00 --end 00:02:30
# Merge multiple videos
uv run scripts/media-process.py merge video1.mp4 video2.mp4 video3.mp4 -o combined.mp4
# Add watermark
uv run scripts/media-process.py watermark input.mp4 --image logo.png --position bottom-right
# Get media info
uv run scripts/media-process.py info input.mp4
AI Generate (Atlas Cloud API) — Gold Sponsor
Generate images and videos using 300+ state-of-the-art AI models. Requires an Atlas Cloud API key.
IMPORTANT for AI agents: Before calling this script, you MUST first use Atlas Cloud MCP tools to find the correct model ID and its required parameters:
- Call
atlas_list_modelsto browse available models, oratlas_search_docsto search for a specific model - Call
atlas_get_model_infowith the model ID to get the exact parameter schema (different models use different parameters — some usesize, others useaspect_ratio+resolution, etc.) - Then call the script with
--model <full_model_id>and the correct parameters
# Generate image (pass full model ID from Atlas Cloud)
uv run scripts/ai-generate.py image "A cat astronaut on the moon" --model black-forest-labs/flux-schnell --size 1024*1024
# Models using aspect_ratio + resolution (e.g. Nano Banana 2, Imagen4)
uv run scripts/ai-generate.py image "Anime girl with blue hair" --model google/nano-banana-2/text-to-image --aspect-ratio 1:1 --resolution 1k
# Models using size presets (e.g. Seedream)
uv run scripts/ai-generate.py image "Product photo on marble" --model bytedance/seedream-v5.0-lite --size 2048*2048
# Edit existing image
uv run scripts/ai-generate.py image "Make the sky sunset orange" --model bytedance/seedream-v5.0-lite/edit --image photo.jpg
# Generate video
uv run scripts/ai-generate.py video "Timelapse of cherry blossoms" --model alibaba/wan-2.6/text-to-video --size 1280*720
# Image-to-video
uv run scripts/ai-generate.py video "The person starts walking" --model alibaba/wan-2.6/image-to-video --image portrait.jpg
# Pass extra model-specific parameters as JSON
uv run scripts/ai-generate.py image "A logo" --model google/imagen4-ultra --extra '{"num_images": 4}'
# NSFW mode
uv run scripts/ai-generate.py image "Artistic figure study" --model black-forest-labs/flux-dev-lora --nsfw
Setup: Set ATLAS_CLOUD_API_KEY in environment or .env file. Get your key at atlascloud.ai.
Output
All tools save output to ./output/ by default. Use -o or --output to specify a custom path.
Models
Models are automatically downloaded on first use and cached locally:
| Tool | Model | Size | Cache Location |
|---|---|---|---|
| Upscale | RealESRGAN_x4plus | ~64MB | ~/.cache/realesrgan/ |
| Face Enhance | GFPGANv1.4 | ~348MB | ~/.cache/gfpgan/ |
| Face Enhance | CodeFormer | ~376MB | ~/.cache/codeformer/ |
| Background Remove | u2net | ~176MB | ~/.u2net/ |
| Object Erase | LaMa | ~200MB | ~/.cache/lama/ |
| Face Swap | buffalo_l + inswapper | ~500MB | ~/.insightface/ |
| Smart Segment | FastSAM-s | ~23MB | auto-downloaded by ultralytics |
Total first-run download: ~1.5GB. All subsequent runs use cached models.
Tips
- GPU Acceleration: All tools automatically use CUDA/MPS if available, falling back to CPU
- Batch Processing: Most tools accept a folder path for batch processing
- Memory: Face swap and segmentation may need 4GB+ RAM for large images
- First Run: First execution downloads AI models — subsequent runs are instant
Workflow Examples
Combine local processing with cloud AI generation:
# 1. Generate a product image with AI
uv run scripts/ai-generate.py image "Minimalist perfume bottle, studio lighting" --model bytedance/seedream-v5.0-lite --size 2048*2048
# 2. Upscale to 4x resolution
uv run scripts/upscale.py ./output/seedream-v5.0-lite_*.png --scale 4
# 3. Remove background for e-commerce
uv run scripts/bg-remove.py ./output/*_x4.png --alpha-matting
# 4. Generate a product video
uv run scripts/ai-generate.py video "A perfume bottle rotating slowly" --model kwaivgi/kling-v3.0-pro/text-to-video --duration 5
# 5. Add watermark to the video
uv run scripts/media-process.py watermark ./output/text-to-video_*.mp4 --image logo.png- 流狐分类
- AI 智能
- 作者声明 Agent
- 未找到明确声明;不据此推断已兼容或已测试
- 静态检查
- 83 / 100 · 启发式扫描,不代表运行安全
- 作者 / 版本 / 许可
- @xramjetx · 未声明 license
- 流狐 Token 估算
- 低消耗
- 流狐接入估算
- 需简单配置
- 是否需要外部 API Key
- 需要 · Vendor-specific
- 检测到的系统要求
- 未声明
- 底层运行要求
- Python >=3.10
- 检测到的文件与系统行为
-
- 只读
- 允许写入 / 修改
- Shell 执行
- 读取环境变量
- 检测到的网络行为
- 允许外网请求
- 安装命令数
- 无(仅作为资料)
档案由构建时根据 SKILL.md 与安装命令自动衍生,可能与作者实际意图存在差异。
需要注意: 未限定 allowed-tools,默认拥有全部工具权限。;检出高风险片段:pipe_curl_to_shell
# Workflow Examples
# 1. Generate a product image with AI
uv run scripts/ai-generate.py image "Minimalist perfume bottle, studio lighting" --model bytedance/seedream-v5.0-lite --size 2048*2048
# 2. Upscale to 4x resolution
uv run scripts/upscale.py ./output/seedream-v5.0-lite_*.png --scale 4
# 3. Remove background for e-commerce
uv run scripts/bg-remove.py ./output/*_x4.png --alpha-matting
# 4. Generate a product video
uv run scripts/ai-generate.py video "A perfume bottle rotating slowly" --model kwaivgi/kling-v3.0-pro/text-to-video --duration 5
# 5. Add watermark to the video
uv run scripts/media-process.py watermark ./output/text-to-video_*.mp4 --image logo.png Python 3.10+ installed uv installed (curl -LsSf https://astral.sh/uv/install.sh | sh) FFmpeg installed (brew install ffmpeg / apt install ffmpeg / winget install ffmpeg)
Tool · Script · What It Does Image Upscale · scripts/upscale.py · 2x/4x super resolution using Real-ESRGAN Face Enhance · scripts/face-enhance.py · Restore and enhance faces using GFPGAN + CodeFormer
All scripts use uv run for zero-setup execution — dependencies are automatically installed on first run.
Upscale low-resolution images by 2x or 4x with AI super resolution.
Restore old photos, enhance blurry faces, fix low-quality portraits.
Remove backgrounds from images, output transparent PNG. Supports multiple AI models.
# Free Image & Video Processing Toolkit
**7 free local AI tools** + **cloud AI generation** (300+ models via Atlas Cloud API).
Local tools run 100% on your machine — no API keys, no cloud costs. Cloud generation tools provide access to state-of-the-art AI models for image and video creation.
## Prerequisites
- Python 3.10+ installed
- [uv](https://docs.astral.sh/uv/getting-started/installation/) installed (`curl -LsSf https://astral.sh/uv/install.sh | sh`)
- FFmpeg installed (`brew install ffmpeg` / `apt install ffmpeg` / `winget install ffmpeg`)
## Available Tools
| Tool | Script | What It Does |
|------|--------|-------------|
| Image Upscale | `scripts/upscale.py` | 2x/4x super resolution using Real-ESRGAN |
| Face Enhance | `scripts/face-enhance.py` | Restore and enhance faces using GFPGAN + CodeFormer |
| Background Remove | `scripts/bg-remove.py` | Remove image backgrounds, output transparent PNG |
| Object Erase | `scripts/erase.py` | Erase unwanted objects using LaMa inpainting |
| Face Swap | `scripts/face-swap.py` | Swap faces between images using InsightFace |
| Smart Segment | `scripts/segment.py` | Segment anything in images using FastSAM |
| Media Process | `scripts/media-process.py` | Convert, compress, resize, extract with FFmpeg |
| **AI Generate** | **`scripts/ai-generate.py`** | **Generate images/videos with 300+ cloud AI models** |
## Usage
All scripts use `uv run` for zero-setup execution — dependencies are automatically installed on first run.
### Image Upscale (Real-ESRGAN)
Upscale low-resolution images by 2x or 4x with AI super resolution.
```bash
# 4x upscale (default)
uv run scripts/upscale.py input.jpg
# 2x upscale
uv run scripts/upscale.py input.jpg --scale 2
# Upscale with face enhancement
… 作者原文负责流程事实;流狐只索引当前章节、要点、文件与命令。
章节 -> Prerequisites → Available Tools → Usage → Image Upscale (Real-ESRGAN) → Face Enhance (GFPGAN + CodeFormer) → Background Remove (rembg)
要点 -> 7 free local AI tools · cloud AI generation · AI Generate · scripts/ai-generate.py · Generate images/videos with 300+ cloud AI models · IMPORTANT for AI agents · Setup · GPU Acceleration
文件/命令 -> curl -LsSf https://astral.sh/uv/install.sh | sh · brew install ffmpeg · apt install ffmpeg · winget install ffmpeg · scripts/upscale.py · scripts/face-enhance.py · scripts/bg-remove.py · scripts/erase.py
内容 SHA-256 -> a1e012ca8e83
方法与流程
适用与边界
原文中的明确线索
curl -LsSf https://astral.sh/uv/install.sh | sh、brew install ffmpeg、apt install ffmpeg、winget install ffmpeg、scripts/upscale.py、scripts/face-enhance.py、scripts/bg-remove.py、scripts/erase.py