本文へスキップ
AI図鑑

切り抜きと背景除去

被写体を背景からきれいに切り出す

画像の生成と編集初級 #22
入力画像画像

本ページの本文は英語で提供されています。タイトルと導入は日本語化されています。

この能力とは何か

Takes an image, sometimes with a hint point or box naming the subject, and outputs an image with an alpha channel — foreground kept, background transparent. Unlike segmentation it outputs a soft alpha matte rather than a hard class mask, and the semi-transparent transition at the edges is exactly what it must get right.

技術的にどう実現するか

A segmentation network first proposes the subject region, then a dedicated matting model estimates per-pixel opacity and foreground colour in the boundary band. Training data comes from compositing known subjects onto random backgrounds plus finely annotated alpha ground truth. Promptable segmentation models have made one-click cutouts easy, and video matting adds a requirement for temporal consistency across frames.

代表的な製品

4

関連する組織

代表的な用途

  • Re-backgrounding and layout of product photos
  • ID-photo and portrait cutouts
  • Virtual backgrounds for short video and live streaming
  • Asset extraction for design and compositing

どう評価するか

IoU
Intersection over union of the foreground region
SAD / MSE
Absolute or squared error of the alpha channel against truth
Gradient error
Whether edge transitions look like a real matte

限界と難しさ

  • Hair, fur and mesh produce ragged or grey fringes instead of clean semi-transparent edges
  • When subject and background colours are close, or the background is busy, the boundary decision fails
  • Frame-by-frame video matting flickers at the edges and needs extra temporal handling

背景にある概念