Concretely, U2ネットセグメンテーションモデルは ピクセルごとのアルファマットを予測し 背景はカットされます しかし損失源は ソフトで圧縮されたエッジを持ちます だからマットは後で 洗練されます 出力は 損失源に戻るよりも アルファチャネルを持つフォーマットに 戻らなければなりません. There is no manual masking step — a neural segmentation model does the separation and you get a transparent result back.
コンバータは,画像をJPEG,PNG,WebP,HEICに移動する。 Deciding the format last rather than first is the whole point: every extra save of a photographic image is another generation, and the fewer of them you spend, the more of the picture survives.