Pith. sign in

REVIEW 18 cited by

FlowAlign: Trajectory-Regularized, Inversion-Free Flow-based Image Editing

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2505.23145 v4 pith:Q6F5PTGB submitted 2025-05-29 cs.CV cs.AIcs.LG

classification cs.CVcs.AIcs.LG
keywords editingflowalignimagesourceconsistentflow-basedinversion-freemethods
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Recent inversion-free, flow-based image editing methods such as FlowEdit leverages a pre-trained noise-to-image flow model such as Stable Diffusion 3, enabling text-driven manipulation by solving an ordinary differential equation (ODE). While the lack of exact latent inversion is a core advantage of these methods, it often results in unstable editing trajectories and poor source consistency. To address this limitation, we propose {\em FlowAlign}, a novel inversion-free flow-based framework for consistent image editing with optimal control-based trajectory control. Specifically, FlowAlign introduces source similarity at the terminal point as a regularization term to promote smoother and more consistent trajectories during the editing process. Notably, our terminal point regularization is shown to explicitly balance semantic alignment with the edit prompt and structural consistency with the source image along the trajectory. Furthermore, FlowAlign naturally supports reverse editing by simply reversing the ODE trajectory, highliting the reversible and consistent nature of the transformation. Extensive experiments demonstrate that FlowAlign outperforms existing methods in both source preservation and editing controllability.

Discussion (0). Sign in to comment.

Forward citations

Cited by 18 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. DirectEdit: Step-Level Accurate Inversion for Flow-Based Image Editing

    cs.CV 2026-05 unverdicted novelty 7.0 of 10

    DirectEdit achieves step-level accurate inversion for flow-based image editing by directly aligning forward paths, using attention feature injection and mask-guided noise blending to balance fidelity and editability w...

  2. FlowAnchor: Stabilizing the Editing Signal for Inversion-Free Video Editing

    cs.CV 2026-04 unverdicted novelty 7.0 of 10

    FlowAnchor stabilizes editing signals in flow-based inversion-free video editing via spatial-aware attention refinement and adaptive magnitude modulation for improved faithfulness and temporal coherence.

  3. Exploring Cross-Modal Flows for Few-Shot Learning

    cs.CV 2025-10 unverdicted novelty 7.0 of 10

    FMA introduces flow matching for multi-step cross-modal feature alignment in few-shot learning, using fixed coupling, noise augmentation, and early-stopping to outperform one-step PEFT methods.

  4. h-Flow: Flexible Flow-based Image Editing via Doob's h-Transform

    cs.CV 2026-07 conditional novelty 6.5 of 10

    h-Flow extends Doob's h-transform to deterministic rectified flows via an equivalent SDE, yielding closed-form reconstruction guidance plus orthogonal velocity editing for controllable text-based image editing.

  5. ElasticTTT: Prior-Preserving Test-Time Tuning for Video Editing

    cs.CV 2026-07 conditional novelty 6.0 of 10

    A test-time tuning framework with three regularization techniques that preserves the generative prior of a video diffusion model during one-shot editing, achieving state-of-the-art results on the authors' benchmark.

  6. ElasticTTT: Prior-Preserving Test-Time Tuning for Video Editing

    cs.CV 2026-07 conditional novelty 6.0 of 10

    Test-time tuning of video diffusion models collapses generation toward the source video; ElasticTTT counters this with noisy targets, contrastive source-prompt guidance, and asynchronous region-wise noise scheduling, ...

  7. Bridging the Manifold Gap: Riemannian Residual Line Search for One-Step Image Editing

    cs.CV 2026-06 unverdicted novelty 6.0 of 10

    Riemannian Residual Line Search improves one-step diffusion image editing by curvature-based residual path construction and CLIP-based candidate selection, reporting SOTA on 700-sample PIE-Bench++ across 10 edit types.

  8. Flow-based Policy Adaptation without Policy Updates

    cs.RO 2026-06 unverdicted novelty 6.0 of 10

    GLOVES learns flow models from limited expert demonstrations to selectively correct actions from non-expert policies or operators toward expert distributions using reverse-flow OOD detection as an intervention gate.

  9. StreamEdit: Training-Free Video Editing via Few-Step Streaming Video Generation

    cs.CV 2026-05 unverdicted novelty 6.0 of 10

    StreamGVE enables high-quality training-free video editing by converting the task to noise-to-data streaming generation with dual-branch fast sampling, self-attention bridges, cross-attention grounding, source-oriente...

  10. StreamEdit: Training-Free Video Editing via Few-Step Streaming Video Generation

    cs.CV 2026-05 unverdicted novelty 6.0 of 10

    StreamEdit enables high-quality training-free video editing by adapting streaming video generation models with dual-branch fast sampling, self-attention bridge, cross-attention grounding, source-oriented guidance, and...

  11. Semantic Granularity Navigation in Image Editing

    cs.CV 2026-05 unverdicted novelty 6.0 of 10

    NaviEdit reallocates fixed step budgets in diffusion rollouts to intermediate scales for improved semantic editability while preserving fidelity via a self-consistency contract.

  12. LimeCross: Context-Conditioned Layered Image Editing with Structural Consistency

    cs.CV 2026-05 unverdicted novelty 6.0 of 10

    LimeCross enables text-guided editing of individual layers in composite images by conditioning on cross-layer context via bi-stream attention while preserving layer integrity and introducing the LayerEditBench benchmark.

  13. Wavelet-Guided Semantic Signal Compensation for Inversion-Free Image Editing

    cs.CV 2026-07 unverdicted novelty 5.0 of 10

    Proposes a frequency-aware semantic compensation strategy using wavelets to strengthen text-conditioned signals in early diffusion steps for better global editing without inversion.

  14. Bridging the Manifold Gap: Riemannian Residual Line Search for One-Step Image Editing

    cs.CV 2026-06 conditional novelty 5.0 of 10

    Second-order curvature-corrected residual line search over energy-field transport candidates yields SOTA one-step text-guided image editing on PIE-Bench++.

  15. DirectEdit: Step-Level Accurate Inversion for Flow-Based Image Editing

    cs.CV 2026-05 unverdicted novelty 5.0 of 10

    DirectEdit eliminates reconstruction error in flow-based image editing by aligning forward paths and applying attention feature injection with mask-guided noise blending.

  16. Translationese as a Rational Response to Translation Task Difficulty

    cs.CL 2026-03 unverdicted novelty 5.0 of 10

    Translationese is partly predictable from quantifiable translation-task difficulty, especially cross-lingual transfer load, more so for English-to-German than the reverse.

  17. FlowSteer: Conditioning Flow Field for Consistent Image Restoration

    eess.IV 2025-12 conditional novelty 5.0 of 10

    A sparse mid-to-late schedule of null-space fidelity updates lets a frozen text-to-image flow model restore images with high measurement consistency.

  18. Semantic Granularity Navigation in Image Editing

    cs.CV 2026-05 unverdicted novelty 4.0 of 10

    NaviEdit is a training-free inference-time controller that decouples edit progress from model scale traversal in diffusion-based image editing via self-consistency, reporting average gains across editors and backbones.

Pith tools