Pith. sign in

Paper Citation Record · LEDGER

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model

As of 19 August 2026, this Paper Citation Record lists 64 of 64 outbound references and 1 inbound Pith citation observation for arXiv:2504.17826.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.17826 v1

Coverage vector

measured 64 of 64 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T10:49:42.241870Z

measured 65 of 65 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T08:24:15.644051Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

64 of 64 outbound references displayed

  • verified exact2
  • verified fuzzy46
  • unresolved16
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d6ebd7ef-2cab-4eda-bdae-30207a101239 · outbound

This paper cites A comprehensive review of circular economy research in the textile and clothing industry,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model A comprehensive review of circular economy research in the textile and clothing industry,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.971514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:41.983098Z digest=sha256:75e3d4f57ec783edfe5c278636b7de2dfdf73a6517e24dfb25a6abd279687a07

Observation fcf3f090-d082-49b4-9eef-9bbd07afcae3 · outbound

This paper cites Hybrid recommender systems: Survey and experiments,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Hybrid recommender systems: Survey and experiments,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.959647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:41.988436Z digest=sha256:954f1c8bad44a03deeb6f1a361d0ffbd79cac4c01c7a6ec0e6c9db7719b8c707

Observation d218ef83-bad9-47f6-8aab-61276da57a93 · outbound

This paper cites The evolution and future of retailing and retailing education,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model The evolution and future of retailing and retailing education,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.947233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:41.992451Z digest=sha256:3317a3cff345fae1b7d0478682baba4967061bce58cbecbf677326cea8ac6226

Observation 9b3de4e0-b6be-4c35-82ad-b2b98c55be52 · outbound

This paper cites Knowledge management and fashion retail performance: the moderating role of product complexity,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Knowledge management and fashion retail performance: the moderating role of product complexity,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.929603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:41.996628Z digest=sha256:a3bfe620dec7256a553e9d48ef9fb5663444373d768819a55f7af12456f98a50

Observation 23df6d15-1792-4a65-935f-0afdd43daaa5 · outbound

This paper cites Assembled or unassembled? different types of outfit coordination presentations in online fashion retailing,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Assembled or unassembled? different types of outfit coordination presentations in online fashion retailing,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.919311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.001584Z digest=sha256:5a701a7dca7d0825da1e6650da845fd5343a4ed03985658ae2acd8616b9a42a0

Observation 5ca171eb-9269-4c97-a562-889fcffe719a · outbound

This paper cites What is the future of fashion retailing with generative ai? understanding consumer response through twitter data,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model What is the future of fashion retailing with generative ai? understanding consumer response through twitter data,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.908261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.005660Z digest=sha256:2365e6233fd3eae92db8feba8aa102a4d8ca9ca164eecad2926aed55a3cac94f

Observation 4ad9fe94-f38e-4dd2-a3c9-4d6bcad1399a · outbound

This paper cites Modeling fashion compat- ibility with explanation by using bidirectional lstm,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Modeling fashion compat- ibility with explanation by using bidirectional lstm,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.897964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.009696Z digest=sha256:49800fa96eaeb28c22818327cd1be17fc464752f0109a39379225901777107ce

Observation d601fc5b-cceb-45db-aaa9-3992d260555b · outbound

This paper cites Fashion forward with ai creations using gan,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Fashion forward with ai creations using gan,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.886767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.013260Z digest=sha256:3698cc7a0bd1a0eb7a956ea12f7b0f64a60f602a1860c0629f807d3b6c31ab64

Observation 9bdba344-d270-45c8-831f-9acbb7fc60a0 · outbound

This paper cites Per- sonalized clothing recommendation fusing the 4-season color system and users’ biological characteristics,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Per- sonalized clothing recommendation fusing the 4-season color system and users’ biological characteristics,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.874852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.016917Z digest=sha256:7d54adbcdaf1375f723e1b92cf25da737ddb35887f524c73c21548e1a997ffe7

Observation 59e376e9-40cc-4bf9-b4ac-aff76ccb13df · outbound

This paper cites The impact of servitization on perceived quality, purchase intentions and recommendation intentions in the ready- to-wear sector,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model The impact of servitization on perceived quality, purchase intentions and recommendation intentions in the ready- to-wear sector,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.864199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.020743Z digest=sha256:01ce88c31bb4674b717461da013b1aeaf130fcf24923f9cd1a83697003a4d139

Observation 38e67e14-3b44-4658-b088-1724cfc4be93 · outbound

This paper cites An intelligent recommendation system in e- commerce using ensemble learning,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model An intelligent recommendation system in e- commerce using ensemble learning,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T10:49:42.024872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:49:42.024872Z digest=sha256:0036444409ec0421fac6fe9f8a8b6632a246e315c66ee9e966a79f6946fb6f81

Observation 49df58c6-66c3-4346-b453-a6c04b04f727 · outbound

This paper cites Learning binary code for personalized fashion recommendation,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Learning binary code for personalized fashion recommendation,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.845495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.029645Z digest=sha256:0946c88be7cba9a71a71fa00775ea3728dabdeed16014b4c6e9fdcf3655fd781

Observation 4223bc97-087a-481f-a5ba-41adcf0c4843 · outbound

This paper cites Learning similarity conditions without explicit supervision,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Learning similarity conditions without explicit supervision,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.834856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.033751Z digest=sha256:b1eeebe2b5e93e830b802fbc5d964640fef7d0b0474f6fcb1e8e2aee7013d620

Observation 03888240-d1d9-4143-8b86-8b4e385d7eb9 · outbound

This paper cites Fashion outfit complementary item retrieval,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Fashion outfit complementary item retrieval,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.823103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.037131Z digest=sha256:e4911bda3a6bd7a4013a43abc1e240aa63ffda6c9f6053a7fb76f7f91dca1d8e

Observation 92a3a12c-c34c-420d-af64-dcc1710a4e3e · outbound

This paper cites Learning type-aware embeddings for fashion compatibil- ity,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Learning type-aware embeddings for fashion compatibil- ity,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.810800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.041190Z digest=sha256:21c1749db3edd241f935c36651cefecdd5ec63c2db317afa13320c336cc1d689

Observation 04f470de-a5e3-48ae-b05e-2d1148969b2a · outbound

This paper cites Language Models are Few-Shot Learners.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Language Models are Few-Shot Learners

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-16T10:49:42.044957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:49:42.044957Z digest=sha256:d5824fdaae6d62ecdb846fb6aed4b6a831dd95c41cfadbf8b20a531a0c803776

Observation 34b984d3-b27b-4c79-a915-5db6fe7bb0d9 · outbound

This paper cites Diffusion models beat gans on image synthesis,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Diffusion models beat gans on image synthesis,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-16T10:49:42.049567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:49:42.049567Z digest=sha256:da28ce7aafabed8378a00b542f81d691038b05fdc8d32fd91bccff6787fe1707

Observation 916039de-4ea0-4f87-8ab9-33d0d13958cc · outbound

This paper cites Outfitgan: Learning compatible items for generative fashion outfits,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Outfitgan: Learning compatible items for generative fashion outfits,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.790837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.053741Z digest=sha256:e3444d300225cf2fd33dd25089375859e64792e12f729045eb3530647f6eb5d1

Observation 35468094-ca76-446a-8c53-f5ee30af0b12 · outbound

This paper cites Diffusion models for generative outfit recommendation,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Diffusion models for generative outfit recommendation,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.779867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.057222Z digest=sha256:9570e8012b7eb6fa70f7c40f7baac095baa4835d5263a0d52d1b8497d6bb3553

Observation b4065064-c0a2-496e-b449-ba67dcf1f27b · outbound

This paper cites Generative Recommendation: Towards Next-generation Recommender Paradigm.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Generative Recommendation: Towards Next-generation Recommender Paradigm

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-16T10:49:42.060893Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:49:42.060893Z digest=sha256:8140720a2e4e83c1c52da9c096fab5fe4d748b5aa8b6de3db909da23017460d0

Observation 7cc38f3c-14b4-4d98-bd98-e1e07b7ba541 · outbound

This paper cites Integrating Domain Knowledge into Large Language Models for Enhanced Fashion Recommendations.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Integrating Domain Knowledge into Large Language Models for Enhanced Fashion Recommendations

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-08-16T10:49:42.393472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.065843Z digest=sha256:892f0d8ffa582a03d86b7c7e7c71445e6317371a4696914040537cfb7aa15ed3

Observation 14a1073a-e31e-421d-891f-8695acba9703 · outbound

This paper cites Fashion recommendation systems, models and methods: A review,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Fashion recommendation systems, models and methods: A review,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.765987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.069789Z digest=sha256:5dd8d6160fc8186721d9504a0a4806f8499a0913afc6f3549ac7deef6338645b

Observation 113184b5-5842-4344-a21c-39fc651835f4 · outbound

This paper cites A review of modern fashion recommender systems,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model A review of modern fashion recommender systems,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-16T10:49:42.073654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:49:42.073654Z digest=sha256:5ad32c1e3d2c3af3089a9f8025204735f03b3c5341e42cdef3c9955d1bec2308

Observation 6921501c-af88-4c7b-9c59-448a31e6dc66 · outbound

This paper cites Generative ai-based style recommendation using fashion item detection and classification,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Generative ai-based style recommendation using fashion item detection and classification,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.753901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.078462Z digest=sha256:371d6246d7f909fda595fb8fa1e7a9d67f8ec8c67551476932b8ab56b9309844

Observation a8c9c96e-8d5b-48ba-a249-86473d7b7951 · outbound

This paper cites Fashioning consumer choices: recommendation, motivation, and purchase intention toward instagram commerce. a mediation analysis,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Fashioning consumer choices: recommendation, motivation, and purchase intention toward instagram commerce. a mediation analysis,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.742186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.082462Z digest=sha256:f23877d1bea4f0c02fc945c0a065e0751a5c4db89f5a4feeb39731248a78b6ae

Observation f1f6d656-2781-4b52-8244-2621037c24ec · outbound

This paper cites Image-based recommendations on styles and substitutes,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Image-based recommendations on styles and substitutes,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-16T10:49:42.087018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:49:42.087018Z digest=sha256:679b1a2ea3a18c708380500c6130a0a6d90e27cba04e1e1eda1b45f2b64351e8

Observation 0bca17aa-a28f-42c9-b5b6-5b88cb44b3c1 · outbound

This paper cites Category-aware mul- timodal attention network for fashion compatibility modeling,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Category-aware mul- timodal attention network for fashion compatibility modeling,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.721727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.090758Z digest=sha256:38f9c05044505cc13609ceb476521b36bb748022b0c3bbddc93da63d54ab8732

Observation 042a6278-ebe6-4b38-978c-22a76d2f86fa · outbound

This paper cites Toward explainable fashion recommenda- tion,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Toward explainable fashion recommenda- tion,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.708638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.094854Z digest=sha256:4ebc204ace9404103a469427f451d071a38c822f1301596ae473295f9464b5e0

Observation a3927eb3-d3e8-4692-8aef-b1b0668c5d92 · outbound

This paper cites Collaborative fashion recommendation: A functional tensor factorization approach,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Collaborative fashion recommendation: A functional tensor factorization approach,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.698118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.098706Z digest=sha256:b7620cd8cee0edf30b0940adeffb7c723e316a5645dae440ed61f2f81d745a04

Observation 1316061f-c01c-4588-9f1f-a089a484a63e · outbound

This paper cites Personalized outfit recom- mendation with learnable anchors,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Personalized outfit recom- mendation with learnable anchors,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.686209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.102562Z digest=sha256:ead46d4909fb492f468514b9d5e3a2d8f255370de033f35e4338177f6af7b9e6

Observation 689557c6-caa8-4125-a7ac-66f34705b0de · outbound

This paper cites Pog: personalized outfit generation for fashion recommendation at alibaba ifashion,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Pog: personalized outfit generation for fashion recommendation at alibaba ifashion,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.675323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.106272Z digest=sha256:7166da0fdd194f7dbee9f152c36070db30f34e16a413f4673c29246331bdea17

Observation 1ad6f48e-fef5-4b09-bca2-e00221bffbd0 · outbound

This paper cites Hierarchical fashion graph network for personalized outfit recommendation,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Hierarchical fashion graph network for personalized outfit recommendation,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.665113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.109441Z digest=sha256:b8d272cdb41085c1f0007e793404bd3ef62898abc48f64d5d2b841363cb684a3

Observation ef82ee6d-951b-4985-b583-5a47653e21aa · outbound

This paper cites Learning visual body-shape-aware embeddings for fashion compatibility,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Learning visual body-shape-aware embeddings for fashion compatibility,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.654748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.113587Z digest=sha256:47dc4cd32204510d2ed357f0cb0359775b9bc2204f6f6b3047f8b53554f4e87b

Observation 5fb99c97-18b9-41f1-b963-e99a32d5fb3a · outbound

This paper cites What dress fits me best? fashion recommendation on the clothing style for personal body shape,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model What dress fits me best? fashion recommendation on the clothing style for personal body shape,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.644253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.117841Z digest=sha256:44222ad713a8be9aa33758f8f8940620e210e72aabc555ce1a4e0a30c23e8053

Observation 29a7a6b5-98b0-4287-bff5-4720dd444ac4 · outbound

This paper cites Hairstyle suggestion using statisti- cal learning,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Hairstyle suggestion using statisti- cal learning,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.634137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.122070Z digest=sha256:7b00891d239bafe7c3d95b429f3f5456bfa779c8cd86ce76f5ba07af200ec4b3

Observation 9cd4eaef-68b1-4fc3-b6d2-2e5aa45f5ecc · outbound

This paper cites Wow! you are so beautiful today!.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Wow! you are so beautiful today!

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.624114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.126593Z digest=sha256:125ec4812a9a305272cb8e7cc609a98bef6904c1ccf90d4ab2c5cf78b5d0befc

Observation 479a73d6-b855-4fb7-a791-04f0a7b11274 · outbound

This paper cites Show me the best outfit for a certain scene: A scene-aware fashion recommender system,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Show me the best outfit for a certain scene: A scene-aware fashion recommender system,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.612449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.130137Z digest=sha256:f3039eddc93d5b11fc78909d898603e8acaaa7907cba5078c79a685805c7a39e

Observation 3ed2e207-a413-427c-b8c7-7d2406d0f19c · outbound

This paper cites A unified framework for outfit design and advice,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model A unified framework for outfit design and advice,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.601950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.134807Z digest=sha256:796d3da232e20428cc8bfd7e1d35fd5ac9d826d85ae1395d17583e38a5b6e567

Observation e3bada71-a438-4903-a925-ac57ad0a463a · outbound

This paper cites Explainable outfit recommendation with joint outfit matching and comment genera- tion,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Explainable outfit recommendation with joint outfit matching and comment genera- tion,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.591247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.138383Z digest=sha256:c54cde294d7aeb830df6d52c8ffe4d2674b4c03a32e5285c7fbaaff6d30c1a93

Observation 646017b4-2335-4d6c-972f-5092d7ea5605 · outbound

This paper cites Fashion outfit generation for e-commerce,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Fashion outfit generation for e-commerce,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.579433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.142267Z digest=sha256:a3d96e1e696b6c43362b33534a421c9c9b6a7af963bfca59775903611836c787

Observation 2a76fc3a-6daa-4be9-a6e9-ef6e2d1dbd83 · outbound

This paper cites Fashion compatibility modeling through a multi-modal try-on-guided scheme,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Fashion compatibility modeling through a multi-modal try-on-guided scheme,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.568817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.145936Z digest=sha256:c30c2e94323609ffd4b222977d93cdc16c6be3d4882a13a50c73929877e3b80a

Observation 5acbc0c6-c13e-49f4-9325-de760b9d5a96 · outbound

This paper cites Knowledge-guided compatibility mod- eling,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Knowledge-guided compatibility mod- eling,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.557797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.149821Z digest=sha256:7d6d10495f56b71d9d0237f13559c28c388fb809a58a5fcfa638386df1e4aae0

Observation c0981c9f-6b8b-468d-8b05-8fca8f371a20 · outbound

This paper cites Outfittransformer: Outfit representations for fashion rec- ommendation,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Outfittransformer: Outfit representations for fashion rec- ommendation,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.546631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.154104Z digest=sha256:440a0bc5fe42d1a30119ad6a77722bd06998af8dc9d67803f35881e4e912ae10

Observation fa579f24-0b81-4238-9cc0-a69009d914fc · outbound

This paper cites Leveraging multimodal features and item-level user feedback for bundle construc- tion,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Leveraging multimodal features and item-level user feedback for bundle construc- tion,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.535235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.158153Z digest=sha256:813f76d1f5357cfd36e0cb7ed2e9862ee316512c560c7684137976f3a9d51c9e

Observation 376f3bef-d490-4731-9cc0-b7ef31a06180 · outbound

This paper cites AI-Yo: Embedding Psychosocial Aspects In the Fashion Stylist Chatbot Design,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model AI-Yo: Embedding Psychosocial Aspects In the Fashion Stylist Chatbot Design,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.523104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.162563Z digest=sha256:184405ea9dc5c7a6f174ec881430d9e6c8c9c3ccdd8f87bd3225890aad56fab8

Observation d412ecd8-1a2f-4cd2-a6c0-36e51fbaf25f · outbound

This paper cites Multimodal conversational fashion recommendation with positive and negative natural-language feedback,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Multimodal conversational fashion recommendation with positive and negative natural-language feedback,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.510418Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.166391Z digest=sha256:d72194cc0bbd90ff6a9eb93f1753afab0caa762a73e25b956226f5370bc0851c

Observation 7a8317ee-331e-46d2-958b-9c597e62340b · outbound

This paper cites Multi-modal dialog state tracking for interactive fashion recom- mendation,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Multi-modal dialog state tracking for interactive fashion recom- mendation,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.499440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.170570Z digest=sha256:0e66f552dcb4a810508ebd705e41481e675a661ced12e072b5e1a67c39249d2a

Observation 0a18937b-87f2-4f53-9a82-2c1d43f8067c · outbound

This paper cites Show-o: One Single Transformer to Unify Multimodal Understanding and Generation.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Show-o: One Single Transformer to Unify Multimodal Understanding and Generation

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-16T10:49:42.174850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:49:42.174850Z digest=sha256:93b286447d908c152b089a9a26d6555cdb20c57c0b4bca4f73bfdfd0e9c5c867

Observation a3550611-3e6f-42f4-ad15-0d675fd31e71 · outbound

This paper cites JanusFlow: Harmonizing Autoregression and Rectified Flow for Unified Multimodal Understanding and Generation.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model JanusFlow: Harmonizing Autoregression and Rectified Flow for Unified Multimodal Understanding and Generation

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-16T10:49:42.179374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:49:42.179374Z digest=sha256:ea09790bc158e6c44bef00753497388d8ee4eaa2feebf19666436589cc015d89

Observation 2de4414d-15ec-4860-9b09-af3631469b4f · outbound

This paper cites Vitron: A Unified Pixel-level Vision LLM for Understanding, Generating, Segmenting, Editing.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Vitron: A Unified Pixel-level Vision LLM for Understanding, Generating, Segmenting, Editing

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-16T10:49:42.183744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:49:42.183744Z digest=sha256:57dab74d73aa9677a43144d53b62c7e9e2a4798a1b06ccb3e0fe56c4ad1b5349

Observation 8023ea8e-1476-45c1-911b-9da72d5165bc · outbound

This paper cites UniFashion: A Unified Vision-Language Model for Multimodal Fashion Retrieval and Generation.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model UniFashion: A Unified Vision-Language Model for Multimodal Fashion Retrieval and Generation

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-16T10:49:42.188238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:49:42.188238Z digest=sha256:22ab8a189ec63323838685da90e03ecc3bbd80ebb10588decff377124749e1a1

Observation d103c76e-a62f-449e-bcc4-9db9d62144dd · outbound

This paper cites Fashionai: A hierarchical dataset for fashion understanding,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Fashionai: A hierarchical dataset for fashion understanding,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.488566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.192936Z digest=sha256:c177b2f4f1743cfb416839c24f17a456946970e562ad5389f77a2874a3534609

Observation 32fbaedb-b21c-4dc0-90ea-ee853f37cc24 · outbound

This paper cites Theme-Matters: Fashion Compatibility Learning via Theme Attention.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Theme-Matters: Fashion Compatibility Learning via Theme Attention

Reference 53

Resolution
verified exact
local_arxiv, observed 2026-08-16T10:49:42.335894Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.197024Z digest=sha256:c406dfe1185e54c5d37f7cc1e36a7c3c583fe1e5ca7c3a773a277e21b334067e

Observation de069976-6d0f-4a81-ad7a-e0880fa44380 · outbound

This paper cites Learning transferable visual models from natural language supervision,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Learning transferable visual models from natural language supervision,

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-16T10:49:42.201520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:49:42.201520Z digest=sha256:082a4824530be311e5573556c67e0ad600dd179e9661879c1559b5ddebc823a0

Observation dd45b2ed-729f-4f7c-b847-8dc746814640 · outbound

This paper cites Language Model Beats Diffusion -- Tokenizer is Key to Visual Generation.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Language Model Beats Diffusion -- Tokenizer is Key to Visual Generation

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-16T10:49:42.204933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:49:42.204933Z digest=sha256:dc237ea95ace68ccfcadf461481fb832128cf23db5f630c983a5418b197395c3

Observation 32f8db0a-dd76-4981-8c6a-8124608e2e13 · outbound

This paper cites Textbooks Are All You Need.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Textbooks Are All You Need

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-16T10:49:42.209513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:49:42.209513Z digest=sha256:47bfb694f8b83e6813b9137ea2c9148c3284df48e9208eac78fe95ab9a0b1520

Observation ea24a450-8af1-4f57-a406-babb0debca26 · outbound

This paper cites Chainlit,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Chainlit,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.469884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.213663Z digest=sha256:4ad1b0ed818641fbdf290c85e6e5fb805f357f8843ae5ff4eacb12fdd2218b10

Observation 7c19551d-caf2-4862-adef-aef3707635c8 · outbound

This paper cites Llama-3.2-11B-Vision,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Llama-3.2-11B-Vision,

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.458673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.217508Z digest=sha256:a559fbdb17d545e9d6e597d24b2ee0071cf09995e63b4033c1192c6d5ae96b86

Observation 8abe641f-b175-4ce7-8758-0a6338a3f0b2 · outbound

This paper cites GPT-4o System Card.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model GPT-4o System Card

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-16T10:49:42.221737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:49:42.221737Z digest=sha256:9261cfc032a3e9f9fdfa173e6524bee0f5e43df11cf5ca458fb786544d71a2b8

Observation e2f3e9d6-9200-4b0a-a4ed-f25704b398a7 · outbound

This paper cites MEDEC: A Benchmark for Medical Error Detection and Correction in Clinical Notes.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model MEDEC: A Benchmark for Medical Error Detection and Correction in Clinical Notes

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-16T10:49:42.226577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:49:42.226577Z digest=sha256:c525fe8246560c4b464cae3b52990e55a668c5c23f879003a1e61becdf7ed2f2

Observation b4f98a40-f1da-43ec-848b-1815959cf879 · outbound

This paper cites Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-16T10:49:42.230853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:49:42.230853Z digest=sha256:2f2035a2211b4ea9a4ed257366c95060d5c511249d15dda2da8aabef08f803ed

Observation 3bb3c7ef-01f5-4ad5-9a97-bfb8c3976ac6 · outbound

This paper cites messages=[{‘role’: ‘system’, ‘content’: ‘‘‘ As a fashion expert, generate a user−system conversation for training a fashion stylist model.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model messages=[{‘role’: ‘system’, ‘content’: ‘‘‘ As a fashion expert, generate a user−system conversation for training a fashion stylist model

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.447897Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.234869Z digest=sha256:5a7af6fc9f33c4e48170cf60dc00913da243fdf78e8afc233c77fc1f798d14eb

Observation 95936c60-9f34-4c63-b94e-17a5a97b8363 · outbound

This paper cites messages=[{‘role’: ‘system’, ‘content’: ‘‘‘ Create a user−system conversation for training a personalized fashion stylist model.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model messages=[{‘role’: ‘system’, ‘content’: ‘‘‘ Create a user−system conversation for training a personalized fashion stylist model

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.436927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.238322Z digest=sha256:2e91610fda71836f10fad3029c32700043c0f75cfc27f17e4a4695a904d7dc4d

Observation f1c15865-64ea-40eb-ab7c-14e9c73bd39e · outbound

This paper cites UN GAYVOE MVEREI LELS.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model UN GAYVOE MVEREI LELS

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.426483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:49:42.241870Z digest=sha256:010b8b1cb369b17a17dd4c08092c01a480154162727b71ef347364c2b1cae8f9

Pith citing papers

Observation 76cc31d9-f9ee-4b35-ab27-39f79bf088c2 · inbound

VOGUE: A Multimodal Dataset for Conversational Recommendation in Fashion cites this paper.

VOGUE: A Multimodal Dataset for Conversational Recommendation in Fashion FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T08:24:15.644051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:24:15.644051Z digest=sha256:0bc8ccb8243d5c6aa7911470a09ec16eb023300a6e3f3d364ad49979774ddf79