Pith. sign in

Paper Citation Record · LEDGER

Scene-based Factored Attention for Image Captioning

As of 15 August 2026, this Paper Citation Record lists 56 of 56 outbound references and 0 inbound Pith citation observations for arXiv:1908.02632.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1908.02632 v3

Coverage vector

measured 56 of 56 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-14T14:44:00.120823Z

measured 56 of 56 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

56 of 56 outbound references displayed

  • verified exact0
  • verified fuzzy44
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 315c9eb5-a4db-463b-a410-1dd72a310c35 · outbound

This paper cites Spice: Semantic propositional image cap- tion evaluation.

Scene-based Factored Attention for Image Captioning Spice: Semantic propositional image cap- tion evaluation

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.925937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:43:59.827858Z digest=sha256:9763917f209c036ba0f58f2f468260f3e1c065c9b3411563d644aab29f78cee9

Observation 4c796bfc-7a31-4222-9fce-82b3fac8c542 · outbound

This paper cites Bottom-up and top-down attention for image captioning and visual question answering.

Scene-based Factored Attention for Image Captioning Bottom-up and top-down attention for image captioning and visual question answering

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.911590Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:43:59.833279Z digest=sha256:64f02a21f0ae318ddf08a341bd6acc925cd218c59d506fff764cb5eddf362c87

Observation 7c170da0-25f9-4254-b73c-bebaaefd668d · outbound

This paper cites Neural Machine Translation by Jointly Learning to Align and Translate.

Scene-based Factored Attention for Image Captioning Neural Machine Translation by Jointly Learning to Align and Translate

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-14T14:43:59.838299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:43:59.838299Z digest=sha256:d060ce7f11c178b2bb54f708e527344465d300ac8d57ad2d57641e685efa05b4

Observation 8302efd5-2b26-45ae-9adf-22ce001be4de · outbound

This paper cites Structcap: Structured semantic embedding for image captioning.

Scene-based Factored Attention for Image Captioning Structcap: Structured semantic embedding for image captioning

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.898623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:43:59.843391Z digest=sha256:c3fb8ceaf293d7535368b0db0d80b3e8b77b7d55836b212670e84f1150829798

Observation 11f9952c-5fdb-4d9c-94c3-90efbac243bb · outbound

This paper cites Groupcap: Group-based image captioning with structured relevance and diversity constraints.

Scene-based Factored Attention for Image Captioning Groupcap: Group-based image captioning with structured relevance and diversity constraints

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.885076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:43:59.848627Z digest=sha256:af348615c4a3281036767822177dd38df9efbe3dee3547c42a7422bd1090fef6

Observation 28931702-8439-443e-bc14-854d8b36f956 · outbound

This paper cites Boosted attention: Leveraging hu- man attention for image captioning.

Scene-based Factored Attention for Image Captioning Boosted attention: Leveraging hu- man attention for image captioning

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.873551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:43:59.854963Z digest=sha256:86cd1bad5e9c05f692b291fa4aa1dfe634da1d566d7086646f3e3ba70b206115

Observation c5732c48-f89c-4aee-ae4d-69fdae7db894 · outbound

This paper cites Show, adapt and tell: Adversarial training of cross-domain image cap- tioner.

Scene-based Factored Attention for Image Captioning Show, adapt and tell: Adversarial training of cross-domain image cap- tioner

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.862258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:43:59.860617Z digest=sha256:8ca269f77d8af1d32413ab40f9fb1d0412d381b2c61a0e3b69f894dc86be505c

Observation ec088722-61bc-4c31-8a9c-a0faf5eecd43 · outbound

This paper cites Microsoft COCO Captions: Data Collection and Evaluation Server.

Scene-based Factored Attention for Image Captioning Microsoft COCO Captions: Data Collection and Evaluation Server

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-14T14:43:59.865303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:43:59.865303Z digest=sha256:0fcf8be1499af7030a6e2f389de2702574255d38844b13b605a86a810f68a7a8

Observation 73010fc5-74a8-4b34-b8ae-d6084e50c023 · outbound

This paper cites Mind’s eye: A recur- rent visual representation for image caption generation.

Scene-based Factored Attention for Image Captioning Mind’s eye: A recur- rent visual representation for image caption generation

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.849822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:43:59.870298Z digest=sha256:47b79d5f1dc5d1738e5393711365409778f277e6d93b7db0a03b0c458d61390e

Observation de26f60e-42c1-46ef-bccc-60a62291b295 · outbound

This paper cites Regularizing rnns for caption generation by reconstruct- ing the past with the present.

Scene-based Factored Attention for Image Captioning Regularizing rnns for caption generation by reconstruct- ing the past with the present

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.835861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:43:59.873861Z digest=sha256:334f13cdbbfc3c41e624ff123991ffe511bb3128c0fcd9357f520996e9907002

Observation 8d0bef78-0fd0-4d34-b357-f707a0520a18 · outbound

This paper cites To- wards diverse and natural image descriptions via a condi- tional gan.

Scene-based Factored Attention for Image Captioning To- wards diverse and natural image descriptions via a condi- tional gan

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.820381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:43:59.878495Z digest=sha256:02962f38c736c6a6e2c6ac6af90c446a303fafa9699a9b4d4f0d91fc51e47f91

Observation d80ce108-32f6-447a-9096-510627b788ef · outbound

This paper cites Meteor universal: Lan- guage specific translation evaluation for any target language.

Scene-based Factored Attention for Image Captioning Meteor universal: Lan- guage specific translation evaluation for any target language

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.806728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:43:59.882738Z digest=sha256:696c88c72752133ca7249a85969994a395785c47798f36f613f264e5c8e6f143

Observation ab4769fe-1639-40fb-8131-d6c2cd63b3ca · outbound

This paper cites Aligning where to see and what to tell: Image cap- tioning with region-based attention and scene-specific con- texts.

Scene-based Factored Attention for Image Captioning Aligning where to see and what to tell: Image cap- tioning with region-based attention and scene-specific con- texts

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.792389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:43:59.887530Z digest=sha256:a85bc6178d5f4c6d242a03c2c33d0c6a45be7ecee1152e40855effa7b1f6a3ed

Observation 894ee86b-a40e-42c6-a1bb-83309987b465 · outbound

This paper cites Stylenet: Generating attractive visual captions with styles.

Scene-based Factored Attention for Image Captioning Stylenet: Generating attractive visual captions with styles

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.778316Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:43:59.896138Z digest=sha256:e59abc9882cfd422ec02ecb0ea1182c631b3f86d325e48511d3221b765138ede

Observation fb003172-98a7-48c8-b6c6-1947828eaf81 · outbound

This paper cites Semantic compositional networks for visual captioning.

Scene-based Factored Attention for Image Captioning Semantic compositional networks for visual captioning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-14T14:43:59.901252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:43:59.901252Z digest=sha256:442908f554caeb2e81723b09016e86410981c72ad32d8d88002abfc5131e10a1

Observation 6a314a0e-15ad-4d44-a47a-3d6c08e8f0ff · outbound

This paper cites Deliberate attention networks for image captioning.

Scene-based Factored Attention for Image Captioning Deliberate attention networks for image captioning

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.753650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:43:59.906665Z digest=sha256:9eb8d9c9e662cade9571230940ca35b93bf765e4f86483e5ca5931914e433221

Observation 903fd968-ac1a-4e39-92bb-15251428f64a · outbound

This paper cites Locally supervised deep hybrid model for scene recogni- tion.

Scene-based Factored Attention for Image Captioning Locally supervised deep hybrid model for scene recogni- tion

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.739388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:43:59.912060Z digest=sha256:bcd8107ba28951549430ef50e5c820475bd6539a12ff95bc7316e76cbfb85cfc

Observation a5f216c4-5e93-4e28-81f7-e309b8efdef8 · outbound

This paper cites Deep residual learning for image recognition.

Scene-based Factored Attention for Image Captioning Deep residual learning for image recognition

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-14T14:43:59.917701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:43:59.917701Z digest=sha256:32f2d9ff160ab1c7bd670a0a480b7a75668b668877350dc48e8a388da47fe533

Observation 47f946bf-bb62-4404-802d-64965fc1da3c · outbound

This paper cites Maximum expected bleu train- ing of phrase and lexicon translation models.

Scene-based Factored Attention for Image Captioning Maximum expected bleu train- ing of phrase and lexicon translation models

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.714749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:43:59.922086Z digest=sha256:d828a48c8a4f9241db1f605baebb1a4509ff0ffa12c06d96a04aea7584aebbb8

Observation 2068a202-762f-4438-aac6-14a2217c8178 · outbound

This paper cites Long short-term memory.

Scene-based Factored Attention for Image Captioning Long short-term memory

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-14T14:43:59.927468Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:43:59.927468Z digest=sha256:1dba1d05c9446f5ed89947384b528e25bb315abb5af0efd29fe360d5c242de19

Observation c0adb881-fceb-4e72-a8ca-46b059bb8705 · outbound

This paper cites Learning to guide decoding for image captioning.

Scene-based Factored Attention for Image Captioning Learning to guide decoding for image captioning

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.692372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:43:59.932880Z digest=sha256:afa5580fe3aa5858132fa3101eaf76cb0f1b4271ed9fb2a1db9379537b3514c5

Observation b7edca00-58c2-427a-a1e9-7e7da4aea4b4 · outbound

This paper cites Deep visual-semantic align- ments for generating image descriptions.

Scene-based Factored Attention for Image Captioning Deep visual-semantic align- ments for generating image descriptions

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.675233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:43:59.937356Z digest=sha256:90eafc700b29024ed65269b7c4d8e6860b0043ff70f0540bb371d9bd8a6573ce

Observation 4023d798-30e8-4f66-847a-b8b68d40a9e7 · outbound

This paper cites Adam: A method for stochastic optimization.

Scene-based Factored Attention for Image Captioning Adam: A method for stochastic optimization

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-14T14:43:59.941775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:43:59.941775Z digest=sha256:ca4a54a7ac8170f7d3dcda7902c611e5599ef5b3f267fbbf35182d2b711e44eb

Observation 45a7c3aa-5357-4cc6-a3dc-23879dd0b8ed · outbound

This paper cites Multi- modal neural language models.

Scene-based Factored Attention for Image Captioning Multi- modal neural language models

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.650770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:43:59.945590Z digest=sha256:6096ef21acd5307c2a1d9ddb1cabaaadae1a3832b719180e4ede6a68d76ac1bd

Observation 1d982e61-1b3d-42b3-b5e3-ded59602f5f2 · outbound

This paper cites Learn- ing hierarchical semantic description via mixed-norm reg- ularization for image understanding.

Scene-based Factored Attention for Image Captioning Learn- ing hierarchical semantic description via mixed-norm reg- ularization for image understanding

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.638434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:43:59.950561Z digest=sha256:ed848c8dd6edc210dfe97822042f9f76cfc4929f12bbd5cc01d3d9141cd87830

Observation bb17c0d5-d29b-4fba-9f8a-477b9165aae2 · outbound

This paper cites Object bank: A high-level image representation for scene classification & semantic feature sparsification.

Scene-based Factored Attention for Image Captioning Object bank: A high-level image representation for scene classification & semantic feature sparsification

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.624939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:43:59.955051Z digest=sha256:3c8af220eaa8c95ab228077973b645381a89b930af2112cb7d720ef0e33e4300

Observation d85e45c8-fbe4-401f-8016-9d8e7d902294 · outbound

This paper cites Rouge: A package for automatic evaluation of summaries.

Scene-based Factored Attention for Image Captioning Rouge: A package for automatic evaluation of summaries

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-14T14:43:59.958905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:43:59.958905Z digest=sha256:3d92ada55d57c851e38c99a15ebe2494aa1a2767294356a8d6b2fbc39b21b76d

Observation 323625d0-acba-4f12-a48b-9046a6247670 · outbound

This paper cites Improved image captioning via policy gra- dient optimization of spider.

Scene-based Factored Attention for Image Captioning Improved image captioning via policy gra- dient optimization of spider

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.602430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:43:59.963373Z digest=sha256:298f6b2a8b6fd435f3f9aee75116d788356b669aacc3e3575a9d91939f9308a2

Observation 86312d41-2dd7-4e74-95df-7341d7ad14a8 · outbound

This paper cites Knowing when to look: Adaptive attention via a visual sen- tinel for image captioning.

Scene-based Factored Attention for Image Captioning Knowing when to look: Adaptive attention via a visual sen- tinel for image captioning

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.586188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:43:59.967522Z digest=sha256:6d754dcfe67ae194ecc6ad702712dca08f0b8fdef8570893ffe4966775c60936

Observation 59698848-7914-4b36-a342-0bbb825678c9 · outbound

This paper cites Discriminability objective for training de- scriptive captions.

Scene-based Factored Attention for Image Captioning Discriminability objective for training de- scriptive captions

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.571266Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:43:59.976422Z digest=sha256:747e7d6825082e03f854ac9aaa580774f68863c2636b317e0dc80895ca4e7612

Observation 1fea4764-3dc6-4b55-8cd0-cb9487344e28 · outbound

This paper cites Deep Captioning with Multimodal Recurrent Neural Networks (m-RNN).

Scene-based Factored Attention for Image Captioning Deep Captioning with Multimodal Recurrent Neural Networks (m-RNN)

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-14T14:43:59.980770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:43:59.980770Z digest=sha256:2d0f5071479e8210378717e48b829c4d7579c575a18adc3dc1b493aedc567438

Observation 2c4d372b-291a-4c5f-bd16-b65cddb81ac5 · outbound

This paper cites Unsupervised learning of image transformations.

Scene-based Factored Attention for Image Captioning Unsupervised learning of image transformations

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.557827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:43:59.985466Z digest=sha256:6721e688a3914cd79eedf1a1df9730944168d10e3d825ce2daf420ac5a6ee192

Observation eea89e18-739d-4c9d-b1be-0211af1a45f7 · outbound

This paper cites Modeling the shape of the scene: A holistic representation of the spatial envelope.

Scene-based Factored Attention for Image Captioning Modeling the shape of the scene: A holistic representation of the spatial envelope

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.542672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:43:59.989637Z digest=sha256:8836a3976ec001ad4a66f6bf0117aeccc171e5a8161050122fc93f6cc907bafa

Observation 79f34f7d-8a43-4a7d-b258-905845a69ab1 · outbound

This paper cites Bleu: a method for automatic evaluation of machine translation.

Scene-based Factored Attention for Image Captioning Bleu: a method for automatic evaluation of machine translation

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.530296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:43:59.993538Z digest=sha256:4d5bd9523520b945a692b48aee34290d9becbf3a7e714ac859553f25dee403c6

Observation ef208dea-a96d-47b9-b49a-5dfe9d4e474e · outbound

This paper cites Relative attributes.

Scene-based Factored Attention for Image Captioning Relative attributes

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-14T14:43:59.997490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:43:59.997490Z digest=sha256:ca1e52c315e2fdd4faae21c50e7ab26419f9e45f315ab60bc824c2a31841b05c

Observation b2d753f4-9897-43b9-a7db-3a4dab7078f8 · outbound

This paper cites Faster r-cnn: Towards real-time object detection with region proposal networks.

Scene-based Factored Attention for Image Captioning Faster r-cnn: Towards real-time object detection with region proposal networks

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.509401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:44:00.003050Z digest=sha256:1e6f9c7202210343fb13f4c92d3c528b5f68b080516df18d21e4e9a1028d73b3

Observation bc5bd7f7-23f2-4137-a349-a6166c4fbc5e · outbound

This paper cites Self-critical sequence training for image captioning.

Scene-based Factored Attention for Image Captioning Self-critical sequence training for image captioning

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.494562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:44:00.013877Z digest=sha256:ea14f37d2f7fb6daaabe645de324bb42629cbcfc658e4022265069b8e63a2a0b

Observation eca88d9e-ac91-4a2f-81a2-721d76014d08 · outbound

This paper cites Biologically inspired fea- ture manifold for scene classification.

Scene-based Factored Attention for Image Captioning Biologically inspired fea- ture manifold for scene classification

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.481364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:44:00.018518Z digest=sha256:f07f796735dbedb6951faa8ffe858400bfd4a6bd4f63bcc41f17bda62ede97da

Observation 49355caa-5f35-4b7a-8379-ca24273a7cd5 · outbound

This paper cites Factored tem- poral sigmoid belief networks for sequence learning.

Scene-based Factored Attention for Image Captioning Factored tem- poral sigmoid belief networks for sequence learning

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.461712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:44:00.024714Z digest=sha256:f14572d014909f3d360ca8a869fd87d9e6d7f0ff854ff11a8d7e8d726faab67d

Observation 5f58d516-a410-4e32-b34e-30fb3550113c · outbound

This paper cites Gen- erating text with recurrent neural networks.

Scene-based Factored Attention for Image Captioning Gen- erating text with recurrent neural networks

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.447073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:44:00.030368Z digest=sha256:df2795170696c02659fb0846c23f52c58eb1fda9b4aa723b06b2595f719a2de3

Observation a8506475-8cfa-4c07-9530-354456f44689 · outbound

This paper cites Sequence to sequence learning with neural networks.

Scene-based Factored Attention for Image Captioning Sequence to sequence learning with neural networks

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.432836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:44:00.036341Z digest=sha256:639adf57e20974994ee0ca2a65f32fdae1273865dd3e207ac0946e3739b1fc9f

Observation 0a11428a-1ac9-429a-a0c9-699e64fb1b0a · outbound

This paper cites Rethinking the inception archi- tecture for computer vision.

Scene-based Factored Attention for Image Captioning Rethinking the inception archi- tecture for computer vision

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-14T14:44:00.041043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:44:00.041043Z digest=sha256:3589d2c1559c261210b7f7204e92d2f776e10cc06480e943bc3e3db8621bc633

Observation 57362296-9b15-41b2-af4b-3dd98e296344 · outbound

This paper cites Factored con- ditional restricted boltzmann machines for modeling motion style.

Scene-based Factored Attention for Image Captioning Factored con- ditional restricted boltzmann machines for modeling motion style

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.402979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:44:00.046436Z digest=sha256:8cec7697f7465dfb5ebd96b9bc28e6c5760f3b97faba7ef649adbe4e1163f602

Observation 8926f0fa-dea3-47f4-88e9-26cd06a467cd · outbound

This paper cites Cider: Consensus-based image description evalua- tion.

Scene-based Factored Attention for Image Captioning Cider: Consensus-based image description evalua- tion

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-14T14:44:00.050375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:44:00.050375Z digest=sha256:c0915046ecf58f7397f0282989e4dcdd76428ac38574a65e8ac9c73c2d4cf3e6

Observation a822829e-f86e-422c-a38d-98b1d8d2c4eb · outbound

This paper cites Show and tell: A neural image caption gen- erator.

Scene-based Factored Attention for Image Captioning Show and tell: A neural image caption gen- erator

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.377521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:44:00.058987Z digest=sha256:280d1098b15b40b0b18303a3b44c43da72794eab5a706266cd51593829cade7c

Observation 633b4075-8110-4f33-9245-0230690a8ca1 · outbound

This paper cites Show and tell: Lessons learned from the 2015 mscoco image captioning challenge.

Scene-based Factored Attention for Image Captioning Show and tell: Lessons learned from the 2015 mscoco image captioning challenge

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.362382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:44:00.065483Z digest=sha256:46174871db060fb92133aac90fc8be87e51c87f17ea1db2822acc498f202e282

Observation c694e35a-19fc-4ed7-9118-9637551401c7 · outbound

This paper cites Knowledge guided disambiguation for large- scale scene classification with multi-resolution cnns.

Scene-based Factored Attention for Image Captioning Knowledge guided disambiguation for large- scale scene classification with multi-resolution cnns

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.348426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:44:00.071017Z digest=sha256:75df44a4878f40df9f0e6f87e0a3191f33cf93ac27d5412e0b4dc6e91c4a83db

Observation c808fb44-12de-400f-8fd4-28d364985b9c · outbound

This paper cites Semantics- preserving bag-of-words models and applications.

Scene-based Factored Attention for Image Captioning Semantics- preserving bag-of-words models and applications

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.336263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:44:00.078127Z digest=sha256:f0d517011f5fe0f0899c0f47e9d5043b9f1314e407874b6751d8b46e98b0b761

Observation c95a16d7-f7cd-4e5b-ad62-bdaa127ce467 · outbound

This paper cites an unresolved cited work.

Scene-based Factored Attention for Image Captioning Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-14T14:44:00.322642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:44:00.083568Z digest=sha256:82013ae7c8e6f5577ee91d2ab0f584e6da66cf1e78c33ea8467726614abd10d6

Observation e175809b-1262-409d-b712-c5ddd7be3bf9 · outbound

This paper cites On multiplicative integration with recurrent neural networks.

Scene-based Factored Attention for Image Captioning On multiplicative integration with recurrent neural networks

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.306182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:44:00.091354Z digest=sha256:bf3e34bac90c346f3058d567b282709c6736eb0af929d0d8e7e3e91519d01b86

Observation dd80c233-acae-4b92-b943-fd0b9178cb70 · outbound

This paper cites Show, attend and tell: Neural im- age caption generation with visual attention.

Scene-based Factored Attention for Image Captioning Show, attend and tell: Neural im- age caption generation with visual attention

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.291109Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:44:00.096489Z digest=sha256:4cfd38be3b2de11a1e54a93d257920b1753ffbd2e9deb49c506fa50360de0408

Observation a2d57d16-a2a4-434d-85a4-61f164b5d19f · outbound

This paper cites Review networks for caption gen- eration.

Scene-based Factored Attention for Image Captioning Review networks for caption gen- eration

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.274286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:44:00.102542Z digest=sha256:a7dccb4e65282c97caf7bf36612e1e57655e012a00db60fcdf157f9f1a1a113e

Observation 45a37c8f-d7bd-4cb3-a4df-1dded6635766 · outbound

This paper cites Boosting image captioning with attributes.

Scene-based Factored Attention for Image Captioning Boosting image captioning with attributes

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.254099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:44:00.107942Z digest=sha256:c2736660b9467109b88156f7278987001b68064e204c09166d2cbece355eb42e

Observation 2cb6420d-9a7b-4f45-8404-af0f6771933f · outbound

This paper cites Image captioning with semantic attention.

Scene-based Factored Attention for Image Captioning Image captioning with semantic attention

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.237269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:44:00.112997Z digest=sha256:42d4c5f1a859ee9f143c15bf74adf21c0f5b79f39e5ef46b0c073d6722d69a86

Observation 03905dd3-64bf-4c62-9a79-63dae95c4fc9 · outbound

This paper cites Pairwise constraints based multiview features fusion for scene classi- fication.

Scene-based Factored Attention for Image Captioning Pairwise constraints based multiview features fusion for scene classi- fication

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.220107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:44:00.117341Z digest=sha256:3d01f40e915734d85d4bbabf8973a224021c611500e08f1d1a5b27d8426745b1

Observation 5b80f39e-ea5c-4284-b806-9e902b78f72d · outbound

This paper cites Learning deep features for scene recognition using places database.

Scene-based Factored Attention for Image Captioning Learning deep features for scene recognition using places database

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:44:00.205895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T14:44:00.120823Z digest=sha256:f6eb2419faad631876e59d66b586e38d75e93d1d6456f3d9fb152f2dd935854d

Pith citing papers

No inbound Pith citation observations are available.