Pith. sign in

Paper Citation Record · LEDGER

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment

As of 13 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 0 inbound Pith citation observations for arXiv:2412.00306.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.00306 v1

Coverage vector

measured 30 of 30 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T05:35:49.359376Z

measured 30 of 30 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

30 of 30 outbound references displayed

  • verified exact2
  • verified fuzzy5
  • unresolved23
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 45a79a92-a61d-41d8-90b9-e7e618e61fe4 · outbound

This paper cites We train the model with a batch size of 192 and drop the image embedding at a rate of 0.1.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment We train the model with a batch size of 192 and drop the image embedding at a rate of 0.1

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:49.929428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T05:35:49.354097Z digest=sha256:1579c86538736230cfcfcf0fe5fc74fdf2f101585ad7b8d742e18626609f67a3

Observation 9f02b43c-6e70-40d3-a3e8-9d95f5b1e620 · outbound

This paper cites Vision Transformers Need Registers.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment Vision Transformers Need Registers

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.234411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.234411Z digest=sha256:176e44efcdfbb78d19994a9adbcdc9e89c25b01a93084a01457c67c3b1018bda

Observation ca0d8515-820f-4219-99e3-62c4b54ed134 · outbound

This paper cites An Image is Worth One Word: Personalizing Text-to-Image Generation using Textual Inversion.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment An Image is Worth One Word: Personalizing Text-to-Image Generation using Textual Inversion

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.240485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.240485Z digest=sha256:9582390a17e2f39a767ca3e72588472d0ea30caca0e213230f67f1e40a27de85

Observation 09952882-1a03-4fb1-b46e-0d2cd9fd92a2 · outbound

This paper cites SwapAnything: Enabling Arbitrary Object Swapping in Personalized Visual Editing.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment SwapAnything: Enabling Arbitrary Object Swapping in Personalized Visual Editing

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.252814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.252814Z digest=sha256:6937833ba85a3198f753ca6bff6e31c777893b25242b5674b1e24371c979d62c

Observation b5cbb33e-cbc6-4312-834c-4c7e6d372533 · outbound

This paper cites COHO: Context-Sensitive City-Scale Hierarchical Urban Layout Generation.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment COHO: Context-Sensitive City-Scale Hierarchical Urban Layout Generation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.258400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.258400Z digest=sha256:a8413e7750ab826c4fc7662768ae5109dd9fa87cfc86abe6bced9807ea69ed79

Observation ea7f2463-4780-4487-9ebb-11fc2e68efce · outbound

This paper cites Kubrick: Multimodal Agent Collaborations for Synthetic Video Generation.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment Kubrick: Multimodal Agent Collaborations for Synthetic Video Generation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.263716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.263716Z digest=sha256:9cb0a43a1c1a91a663806cddd5ffdcb617e97e131f1e57fbebab7b1fc091fb0b

Observation ecdbac20-165d-4508-b0c1-0b94cdcde5d1 · outbound

This paper cites Prompt-to-Prompt Image Editing with Cross Attention Control.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment Prompt-to-Prompt Image Editing with Cross Attention Control

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.270326Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.270326Z digest=sha256:6661a4c13b8d970edd3ae09af925308a7e49ee32313a60528dfeae767447f919

Observation 69d610d4-e174-4bed-9d6f-423da042533e · outbound

This paper cites Multi-concept customization of text-to-image diffusion.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment Multi-concept customization of text-to-image diffusion

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:49.983476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T05:35:49.281681Z digest=sha256:9fbe3a9d032d2f3654e33093fe08b0ed3d310cc52fe0dc70866e7ed98bcd73ba

Observation dda2c42b-1a2d-4700-a094-67ce62b2702b · outbound

This paper cites BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.287588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.287588Z digest=sha256:7b9c4e59a548783a707b5be25e458394ec911aefaf27c0148da772d14aaf58f8

Observation a337f01a-bfa9-44ef-9e2e-db61bc6bc27e · outbound

This paper cites Unihuman: A unified model for editing human images in the wild.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment Unihuman: A unified model for editing human images in the wild

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:49.964355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T05:35:49.292725Z digest=sha256:1b5941dfec924f3100ec34228395c261abab69fe0de2614bf1e6e687d03ed467

Observation 701b4eaf-ac43-4827-a4f1-f753e81ef7db · outbound

This paper cites CliqueParcel: An Approach For Batching LLM Prompts That Jointly Optimizes Efficiency And Faithfulness.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment CliqueParcel: An Approach For Batching LLM Prompts That Jointly Optimizes Efficiency And Faithfulness

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-12T05:35:49.665549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T05:35:49.297777Z digest=sha256:24f35aa027b457ac0bd6833bd97d22a0bc9db23435317e6ebd2f3f212e168802

Observation 95aa426a-84bb-4ecc-b7f6-f3b4bab6c186 · outbound

This paper cites Latent Consistency Models: Synthesizing High-Resolution Images with Few-Step Inference.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment Latent Consistency Models: Synthesizing High-Resolution Images with Few-Step Inference

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.302850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.302850Z digest=sha256:0b4f36be687da15e55c1f38e63e25739d39ff4f26fcc4006203e3d41575dfd56

Observation b19f6622-8eb3-48cf-869a-8f78ede00d03 · outbound

This paper cites Kosmos-G: Generating Images in Context with Multimodal Large Language Models.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment Kosmos-G: Generating Images in Context with Multimodal Large Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.307584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.307584Z digest=sha256:2df35b21c13d8eb4f548b56778ed47034c65abf86b4ba2b5ea50bcf6c1a6f4b0

Observation 7bd69111-d137-4f0d-8427-83c770c72f45 · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.313063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.313063Z digest=sha256:e6913ac3743c55f83000431f8f36110bf66c4830199ab1ca09ede7dfd8160965

Observation f3218787-7d44-4fad-9acf-36a2f85c2ecb · outbound

This paper cites Orb: An efficient alternative to sift or surf.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment Orb: An efficient alternative to sift or surf

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:49.946537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T05:35:49.317747Z digest=sha256:7db233a2f4f40abce20b11b5c433f53334beddd6c9bd8d71b777e1d8a33cdca7

Observation ea202a3a-e097-472c-b5e6-7492ba2629e4 · outbound

This paper cites InstantBooth: Personalized Text-to-Image Generation without Test-Time Finetuning.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment InstantBooth: Personalized Text-to-Image Generation without Test-Time Finetuning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.327527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.327527Z digest=sha256:c75b2afbe730477fef429083cec26ecc09fc5b9cda244e2275b282433ec15d28

Observation 3593f28f-97c4-45b6-aee3-b6e3716cace5 · outbound

This paper cites Empower- ing llms with pseudo-untrimmed videos for audio-visual temporal understanding, 2024b.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment Empower- ing llms with pseudo-untrimmed videos for audio-visual temporal understanding, 2024b

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.332573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.332573Z digest=sha256:df2f5b5eea9f52e52886166e130fabb3756c4880929868f0e3cdfc1fbb1c26b5

Observation 0af59515-f522-40d6-9c5a-0ab4f95a5bbb · outbound

This paper cites GroundingBooth: Grounding Text-to-Image Customization.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment GroundingBooth: Grounding Text-to-Image Customization

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.338092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.338092Z digest=sha256:634ac366e54e2cd377b3813a575966cb4a1b23b65259a7d2440f8f87665d97a5

Observation a799112a-eb4c-4033-8e53-411439055035 · outbound

This paper cites PromptFix: You Prompt and We Fix the Photo.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment PromptFix: You Prompt and We Fix the Photo

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.343195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.343195Z digest=sha256:76cdd1e4e2fd683e09b03207139216b8ce60d6a6b66cfd152778f52d7c5fa196

Observation c8779f89-33b1-4f36-bef6-f6c49fd042a0 · outbound

This paper cites LLMExplainer: Large Language Model based Bayesian Inference for Graph Explanation Generation.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment LLMExplainer: Large Language Model based Bayesian Inference for Graph Explanation Generation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.348093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.348093Z digest=sha256:aa9554e981ade5cb17dca28319ba916246386035ec21287fd1ff5b46620cab13

Observation 0301323f-f1b7-455c-8eca-c6ceea8c7ff5 · outbound

This paper cites an unresolved cited work.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-12T05:35:49.912416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T05:35:49.359376Z digest=sha256:d366bbb68b32cd049c8c403c4f7e346bc96815f5a1919d3de9681ae75b9bda22

Observation dff6126b-5035-4045-830f-27558ee014a0 · outbound

This paper cites SynArtifact: Classifying and Alleviating Artifacts in Synthetic Images via Vision-Language Model.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment SynArtifact: Classifying and Alleviating Artifacts in Synthetic Images via Vision-Language Model

Reference 2006

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.212116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.212116Z digest=sha256:34d40c37ec98bc042b8ea97ca999f6a4e475e73fdba6d218daa5946efe3df37a

Observation bf1480ec-7946-4984-ac1c-670d7bfe6848 · outbound

This paper cites HyperDreamBooth: HyperNetworks for Fast Personalization of Text-to-Image Models.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment HyperDreamBooth: HyperNetworks for Fast Personalization of Text-to-Image Models

Reference 2011

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.322696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.322696Z digest=sha256:a4aeb71cf6536349511ead8beb849c68b2929365ccc77904b99b15b91d95820c

Observation 85720df2-7b3b-4219-968a-73fb03e7488e · outbound

This paper cites Photoswap: Personalized Subject Swapping in Images.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment Photoswap: Personalized Subject Swapping in Images

Reference 2014

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.246983Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.246983Z digest=sha256:caaecc5511e1d58d1f025dc169dd6ba14cc6f91fb00a2017451e539989d4766f

Observation ea0c4b71-066b-441c-aa6e-0139cb5ccc8a · outbound

This paper cites Cross- image attention for zero-shot appearance transfer.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment Cross- image attention for zero-shot appearance transfer

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.200848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.200848Z digest=sha256:42d27c42f15812a82dd3d0735e1aed7f1294cd3c871ddf96fde002afa8eec325

Observation e95f8177-7c36-4217-9b7c-e29fea831f5c · outbound

This paper cites FINEMATCH: Aspect-based Fine-grained Image and Text Mismatch Detection and Correction.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment FINEMATCH: Aspect-based Fine-grained Image and Text Mismatch Detection and Correction

Reference 2020

Resolution
verified exact
local_arxiv, observed 2026-08-12T05:35:49.707068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T05:35:49.276116Z digest=sha256:4159b52fe192079d0065069b444980d0f798bc1a32bb157523cba53e287160db

Observation 215b6e92-e1ad-44a2-b4fd-42ad8416616f · outbound

This paper cites Improving Diffusion Models for Authentic Virtual Try-on in the Wild.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment Improving Diffusion Models for Authentic Virtual Try-on in the Wild

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.228583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.228583Z digest=sha256:74ea81b3cd98424b4c41b6528bf28ec462b98c5b1d70451821440d29524809c7

Observation 983caa76-1d01-4e5b-9705-87edcc664404 · outbound

This paper cites Surf: Speeded up robust features.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment Surf: Speeded up robust features

Reference 2022

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:50.001997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T05:35:49.207143Z digest=sha256:fbaeaab94f66f1180e87037efb691501c448a960ee156017b28749a33211dc3e

Observation 1cccc3fc-e6d1-4f31-ba60-a96a4f2463c1 · outbound

This paper cites Zero-shot Image Editing with Reference Imitation.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment Zero-shot Image Editing with Reference Imitation

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.222660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.222660Z digest=sha256:7782f674dd73da2d6a45ce052c9806138982176202d9257b45fd3477f0ee62ca

Observation eade9098-e1f0-4e26-8019-b16d720e4f17 · outbound

This paper cites AnyDoor: Zero-shot Object-level Image Customization.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment AnyDoor: Zero-shot Object-level Image Customization

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.217312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.217312Z digest=sha256:7a5cd44ccbbbee11e16e1e7c0d913d04703a0a99e7f86e95fc451f7736b73676

Pith citing papers

No inbound Pith citation observations are available.