Pith. sign in

REVIEW

The Animation Transformer: Visual Correspondence via Segment Matching

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2109.02614 v2 pith:YSJQT7PQ submitted 2021-09-06 cs.CV cs.AIcs.GR

classification cs.CVcs.AIcs.GR
keywords animationcorrespondencevisualbuildingenableshand-drawnimageslearn
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Visual correspondence is a fundamental building block on the way to building assistive tools for hand-drawn animation. However, while a large body of work has focused on learning visual correspondences at the pixel-level, few approaches have emerged to learn correspondence at the level of line enclosures (segments) that naturally occur in hand-drawn animation. Exploiting this structure in animation has numerous benefits: it avoids the intractable memory complexity of attending to individual pixels in high resolution images and enables the use of real-world animation datasets that contain correspondence information at the level of per-segment colors. To that end, we propose the Animation Transformer (AnT) which uses a transformer-based architecture to learn the spatial and visual relationships between segments across a sequence of images. AnT enables practical ML-assisted colorization for professional animation workflows and is publicly accessible as a creative tool in Cadmium.

Discussion (0). Continue with ORCID to comment.

Pith tools