Pith. sign in

Paper Citation Record · LEDGER

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain

As of 7 August 2026, this Paper Citation Record lists 51 of 51 outbound references and 2 inbound Pith citation observations for arXiv:2505.14826.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.14826 v1

Coverage vector

measured 51 of 51 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:33:55.702118Z

measured 53 of 53 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-22T09:19:39.848194Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-22T09:21:21.452805Z

Reference resolution

51 of 51 outbound references displayed

  • verified exact2
  • verified fuzzy24
  • unresolved24
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation dcdd67a7-9a4f-48fe-a4fb-2511c49dbc64 · outbound

This paper cites write newline.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T15:33:49.231262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:33:49.231262Z digest=sha256:d31563020eda94215c644d88f26108f4a77ac4cd1fdd064fbc047cc720ed8df7

Observation 424a77be-2830-4c67-badd-cb43a352dbb6 · outbound

This paper cites SemDeDup: Data-efficient learning at web-scale through semantic deduplication.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain SemDeDup: Data-efficient learning at web-scale through semantic deduplication

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T15:33:49.406344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:33:49.406344Z digest=sha256:06f241a0dffed49de98058c418318f7660f9906cb301f8b3d81f6f6e69a4e23d

Observation b7fa811c-abfe-4f18-866b-5bcee96213db · outbound

This paper cites Improved algorithms for linear stochastic bandits.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain Improved algorithms for linear stochastic bandits

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:34:03.803499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T15:33:49.527948Z digest=sha256:6e184319f0efb9cfc29f1e8154f135c5a1e7e1322fae6b33a708704f6a6db789

Observation 9d299939-1bc1-4683-bf7a-e12260681d37 · outbound

This paper cites P., and Wunder, M.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain P., and Wunder, M

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:34:03.580345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T15:33:49.736418Z digest=sha256:9dd097626724110b905b8f9b7fb1f10d888a9e4eaf0156b86fc3d63c044d93d0

Observation 10456815-691d-4ac8-8d51-eb7b5e19b2cf · outbound

This paper cites an unresolved cited work.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:34:03.320041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T15:33:49.868934Z digest=sha256:d4f9e22ba69108140566e52c43d144f4b4d75091f24437cc2bf93a3ad2a04096

Observation 2b3129c8-81ee-493d-ab6c-ba184d858840 · outbound

This paper cites Pattern Recognition and Machine Learning.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain Pattern Recognition and Machine Learning

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:34:03.036966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T15:33:50.045825Z digest=sha256:d78a03df793564d74317728f27642d091232531b09dd7657afb37319d29aef1f

Observation 440e0be3-022f-4413-ace6-9fee82a33c73 · outbound

This paper cites On the Opportunities and Risks of Foundation Models.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain On the Opportunities and Risks of Foundation Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T15:33:50.215979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:33:50.215979Z digest=sha256:6912cbef7629f286e672bcd5ce373c4f7a525a949186c10d67ea732765d979d2

Observation a2b9c314-426e-4062-9f3f-45f48fb1accc · outbound

This paper cites Coresets via bilevel optimization for continual learning and streaming.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain Coresets via bilevel optimization for continual learning and streaming

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:34:02.789773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T15:33:50.317302Z digest=sha256:e1210408d6bcc5f60b3e97d86ed079c4c8e7aa79f2adecb1d2b164b7974d435f

Observation b21aa86a-3b3f-44c6-bf34-cd3a21d39834 · outbound

This paper cites an unresolved cited work.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:34:02.506276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T15:33:50.422295Z digest=sha256:eedccf4c0683fe71f10c5f0fcec65dad0bc622a2a7455e53a5a74e96a0f18edf

Observation 43649b05-d5d4-46ca-a0c7-d1c2ffd0d77f · outbound

This paper cites Super-Samples from Kernel Herding.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain Super-Samples from Kernel Herding

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T15:33:50.595074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:33:50.595074Z digest=sha256:9f0d2ebef37a766c4a2bf4abf4019ade20d0a7f3c2b0e8d4a6a3272f3a28a31f

Observation a0f747e9-5d66-46c1-b1db-1038e62a893a · outbound

This paper cites M., Haussmann, E., and Fardet, E.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain M., Haussmann, E., and Fardet, E

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:34:02.258315Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T15:33:50.784785Z digest=sha256:96b5548d9bc3451bc0274542adbe4d871cbb1e8f777d2693140553113410ec4c

Observation 14c92296-4c4d-42b5-8064-04250ade4892 · outbound

This paper cites and Shrivastava, A.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain and Shrivastava, A

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T15:33:50.927740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:33:50.927740Z digest=sha256:0cfb0479c398c1865a4897f150fc606e0147aa54bfab64e7c597bb87ebf2fb16

Observation 65ea28c3-3bc3-459f-84d4-69e99929c55e · outbound

This paper cites Selection via proxy: Efficient data selection for deep learning.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain Selection via proxy: Efficient data selection for deep learning

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:34:02.038794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T15:33:51.076260Z digest=sha256:c04a86c00dc6e024e1857a8eeeb30202abb1854e3ed456d3b5b13c25ad614298

Observation fc84104b-b3ff-4fd6-aee8-135c2a4557f2 · outbound

This paper cites Active Preference Optimization for Sample Efficient RLHF.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain Active Preference Optimization for Sample Efficient RLHF

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T15:33:51.144721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:33:51.144721Z digest=sha256:d30255041072f3e7d9cfb8a696ed8cbeb03b1d884fd1043a400dd226c5f494a5

Observation 66056af6-47f9-4bbb-b5bb-6a1465fbfc34 · outbound

This paper cites and Zhang, C.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain and Zhang, C

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:34:01.745072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T15:33:51.262928Z digest=sha256:f7e631ec2ddc6e929d215ee5e0ea420855df03662bec0727204d35cd8f4e34d8

Observation 7de814da-df92-4acb-bb0f-ea1e5cc4ea01 · outbound

This paper cites On the mathematical foundations of theoretical statistics.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain On the mathematical foundations of theoretical statistics

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:34:01.466564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T15:33:51.401001Z digest=sha256:7c9ce684f5a37f76864b98aa04d5594020c45a1163bf50e69e9df19a25d7a4a8

Observation 3b88b5e0-1bfd-4e8f-b11c-5dd4c87ee52b · outbound

This paper cites Minimax-optimal Inference from Partial Rankings.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain Minimax-optimal Inference from Partial Rankings

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:33:56.609259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T15:33:51.500852Z digest=sha256:a228bd408c73ea91355b2792598fcab0146f7c504f5a20495913742e94cbfdda

Observation c516dfb3-4f8d-464c-ad70-4383042b3d25 · outbound

This paper cites an unresolved cited work.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T15:33:51.671297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:33:51.671297Z digest=sha256:2c530293dfa8e9341a6586ab8752e7283a06e5e5ab00e0b8780e0d382b1c01d9

Observation e9950644-7869-4f40-9e56-0798500eed05 · outbound

This paper cites LoRA : Low-rank adaptation of large language models.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain LoRA : Low-rank adaptation of large language models

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:34:01.254404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T15:33:51.777301Z digest=sha256:7456d1433ef24acfbe23bc22bb909e489e7476782706630fbf95e586e826e81b

Observation f930385d-710a-4e19-88fa-ecf1990b1214 · outbound

This paper cites an unresolved cited work.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T15:33:51.858503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:33:51.858503Z digest=sha256:860aae80aa1a126648b97af6d775f59a7aca8122b883a1bb571a771c7eb8e6b6

Observation 20223e23-d6a5-494f-b997-e27c74baff71 · outbound

This paper cites and Liberty, E.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain and Liberty, E

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:34:00.933462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T15:33:51.923206Z digest=sha256:27aa12da1fd0df31c66b8343b549f80142284f6a5245fc0bb5c0e2753feff3be

Observation 288ac579-d70f-4519-a0cf-39f96417ac91 · outbound

This paper cites char-rnn.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain char-rnn

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:34:00.669480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T15:33:52.053201Z digest=sha256:47e458e4d4fb74de2b347eef2eec465341efc306ec51c19a5aa75d9a9f54573a

Observation 3dc3c016-1e45-4bd0-9568-9e1fa1bc6a41 · outbound

This paper cites and Szepesvari, C.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain and Szepesvari, C

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:34:00.446532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T15:33:52.168727Z digest=sha256:2919b39ec8ee8a27087638b0f4de22c38109fd8f7dcb79a628521162febf01a0

Observation c632ca69-546e-4b6a-86ea-8a88854e1ca4 · outbound

This paper cites Beyond Scale: The Diversity Coefficient as a Data Quality Metric for Variability in Natural Language Data.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain Beyond Scale: The Diversity Coefficient as a Data Quality Metric for Variability in Natural Language Data

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T15:33:52.314478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:33:52.314478Z digest=sha256:3d73043d27d618d59319e33ec22b75fd559206f24aac27bcd6a3195fdd180818

Observation c715a6c3-925b-4021-87a9-c8303ad1fe6d · outbound

This paper cites Deduplicating training data makes language models better.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain Deduplicating training data makes language models better

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:34:00.165111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T15:33:52.460469Z digest=sha256:c8a992aa9a85f491ab78bac922c6d3393d346334a528daa580456da46c071ed9

Observation 11e2c045-38c1-4fd7-85f6-88d51aca9b29 · outbound

This paper cites Dual Active Learning for Reinforcement Learning from Human Feedback.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain Dual Active Learning for Reinforcement Learning from Human Feedback

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T15:33:52.622227Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:33:52.622227Z digest=sha256:ca10373b71eab7b4baa630314a0505b63ec2a1ba9ef4b496b32c666e27b51a99

Observation df986964-b692-4a24-92f0-95cb5c07d13a · outbound

This paper cites Peft: State-of-the-art parameter-efficient fine-tuning methods.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain Peft: State-of-the-art parameter-efficient fine-tuning methods

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T15:33:52.795149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:33:52.795149Z digest=sha256:48d779f03fdfb9b68a62cdd3238ca6c38d8a7612b9a7de399caea81dece98121

Observation 79fcef58-9e46-4540-ac9f-6617a2db6e4d · outbound

This paper cites Trivial or impossible -- dichotomous data difficulty masks model differences (on ImageNet and beyond).

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain Trivial or impossible -- dichotomous data difficulty masks model differences (on ImageNet and beyond)

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:33:56.330976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T15:33:52.929278Z digest=sha256:438993aec2c55b9a5d44134c323b17f1a1bb8bee16e767dd088c55b052cc4f8a

Observation b7cd40f1-9e20-4086-b05c-5f5af4492c3e · outbound

This paper cites S., and Dean, J.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain S., and Dean, J

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:33:59.839731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T15:33:53.004777Z digest=sha256:c7e5f1f8b167e202d09423c36cf37b324ba69f162e73a7b58920caa46140a25d

Observation 1e9e8a58-a9c4-46ae-af55-da9d1fdc3dad · outbound

This paper cites Prioritized training on points that are learnable, worth learning, and not yet learnt.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain Prioritized training on points that are learnable, worth learning, and not yet learnt

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:33:59.642934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T15:33:53.114717Z digest=sha256:1dbdbed2991adc4d4d17f005b359902c7baacc35f2a1c2ce66ce7e51850efa3a

Observation 572be791-6b49-425e-bb1c-fa4610de8ccf · outbound

This paper cites The star-shaped space of solutions of the spherical negative perceptron.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain The star-shaped space of solutions of the spherical negative perceptron

Reference 31

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T15:33:56.128852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T15:33:53.224174Z digest=sha256:bb41a39d4112dd183753e89b4632382df183d338e824a4e4694678078370b55f

Observation 4cc89c8e-a620-4d37-8b83-0ff9f53369c6 · outbound

This paper cites Optimal design for human preference elicitation.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain Optimal design for human preference elicitation

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:33:59.435562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T15:33:53.335769Z digest=sha256:ef7bb1ad098b054ff74e24cd8b46d43e21ee8e0a39cd3d302e005ec682b9ce95

Observation de061543-acf5-46c6-9b9d-8f06fc21807a · outbound

This paper cites L., Wolsey, L.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain L., Wolsey, L

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:33:59.171920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T15:33:53.417518Z digest=sha256:d1aaa47952aaa2062a2aab7dcbfadc521d8147362660632834a614bdcd343a2d

Observation 77fca0d3-9b9e-4280-b16c-95e8b1e666e1 · outbound

This paper cites Training language models to follow instructions with human feedback.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain Training language models to follow instructions with human feedback

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:33:58.951173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T15:33:53.541745Z digest=sha256:9036833f3cb15f56dd197c84070d261ddf54cbf46eb2a0e0ba06bf1b3fd5ff06

Observation b6fea7d4-6d13-46f5-8e20-86f44525c2f7 · outbound

This paper cites an unresolved cited work.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:33:58.699653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T15:33:53.646729Z digest=sha256:d6764298ee2a17ee4ebdddd282d63ae3f8accb23844b1f480a500d942f39d53b

Observation 13517ba4-ba3e-442a-81ae-5d12ae33a063 · outbound

This paper cites an unresolved cited work.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:33:58.416602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T15:33:53.761624Z digest=sha256:220af46c466fb8b1c5d19dc1c4b273628daad8cbd2b569935d4ed546a85934c6

Observation 9d6e6a10-bf62-4d2e-99a2-a92d521ce1e2 · outbound

This paper cites Optimal Design of Experiments, volume 50 of Classics in Applied Mathematics.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain Optimal Design of Experiments, volume 50 of Classics in Applied Mathematics

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:33:58.108840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T15:33:53.841150Z digest=sha256:cfd8dffcf374a250654539aa89af6396239e673a413a7686b39a7967739c78a9

Observation 94379257-47b7-4a15-8c47-503f79d20f98 · outbound

This paper cites Language models are unsupervised multitask learners.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain Language models are unsupervised multitask learners

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:33:57.859146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T15:33:53.968315Z digest=sha256:215b26919a5b05b29bdf9769ca59ea5586c616ddf36d7cd0fa64ffa61f60d821

Observation db80c7d6-3f9c-4ad1-8d2e-e5bff9fed775 · outbound

This paper cites Direct preference optimization: Your language model is secretly a reward model.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain Direct preference optimization: Your language model is secretly a reward model

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:33:57.718365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T15:33:54.059317Z digest=sha256:c148d1f25080d3b5797ff43db9f9a880be4c69cb1278bc16f357102ddb86f7f9

Observation 05c8e0f8-fa83-4633-9b7b-aa91d74cacef · outbound

This paper cites SVP-CF: Selection via Proxy for Collaborative Filtering Data.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain SVP-CF: Selection via Proxy for Collaborative Filtering Data

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T15:33:54.228168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:33:54.228168Z digest=sha256:63885255667a36c871e4e7a9ad84605c86efb4b44ee270e874b8b029de19fdfc

Observation 75691a37-0af0-4eb6-8cbb-6578d75e4297 · outbound

This paper cites How to Train Data-Efficient LLMs.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain How to Train Data-Efficient LLMs

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T15:33:54.356737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:33:54.356737Z digest=sha256:2270abb2f0ef37568b64ecf3d5991086e111f0e34b6bce5a4f1d724538649556

Observation c58e1577-1621-4ef2-9eb0-c87452a42a56 · outbound

This paper cites Optimal Design for Reward Modeling in RLHF.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain Optimal Design for Reward Modeling in RLHF

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T15:33:54.491558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:33:54.491558Z digest=sha256:3f3f505abb18451627260abef9ce5cfaa9a03dd343e83595b1c193530b091fef

Observation f9715298-d517-417e-b192-2b2340faf1e2 · outbound

This paper cites an unresolved cited work.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:33:57.554704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T15:33:54.652169Z digest=sha256:4da1047fce7be42095a0c18eab4cc857682b87b3f66d8c82f3148d311455948b

Observation 16d7da80-b2c7-4f87-8fec-eb8b86ae27cc · outbound

This paper cites and Yang, M.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain and Yang, M

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:33:57.419423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T15:33:54.730925Z digest=sha256:9b8d07a60f5f7116584d62553bab18373d47c1e2859b858abe8e762036c80052

Observation 64413cef-0d91-4807-9ca8-f8e52ee96539 · outbound

This paper cites Comparing Few to Rank Many: Active Human Preference Learning using Randomized Frank-Wolfe.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain Comparing Few to Rank Many: Active Human Preference Learning using Randomized Frank-Wolfe

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T15:33:54.880840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:33:54.880840Z digest=sha256:9eae348ecc6b5874f3e1ea43ce0e680cd5bb012f809bc568844d90274da4fe67

Observation cc034977-d9fb-489d-9ff7-ad39c71ea6ce · outbound

This paper cites D4: Improving LLM Pretraining via Document De-Duplication and Diversification.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain D4: Improving LLM Pretraining via Document De-Duplication and Diversification

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T15:33:55.009360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:33:55.009360Z digest=sha256:d88707c782896344d14127eda20c44f47aa8020f27ddb32bf59f4e842344385d

Observation f0bc88ca-ac6d-4833-9c01-9e5f391100cd · outbound

This paper cites On coresets for support vector machines.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain On coresets for support vector machines

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:33:57.238907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T15:33:55.132746Z digest=sha256:911f97d0842e9c54dbf75fa9da989fb6a070aed900fa9903d591976560b16d86

Observation 0b35729e-5db2-4b13-8ab3-097ecd82dbef · outbound

This paper cites W., Lester, B., Du, N., Dai, A., and Le, Q.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain W., Lester, B., Du, N., Dai, A., and Le, Q

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:33:57.000752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T15:33:55.280820Z digest=sha256:d41d16fa2cf59be34478382914299618aaf65a74940fb2f3b554ad0c544e7ba3

Observation 253f9be1-f33b-44e4-8048-6bebdc1a3247 · outbound

This paper cites CCNet: Extracting High Quality Monolingual Datasets from Web Crawl Data.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain CCNet: Extracting High Quality Monolingual Datasets from Web Crawl Data

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T15:33:55.441839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:33:55.441839Z digest=sha256:21b7e0480184cf2e30acb8d12ab8cda41f05c84a9a374164f9bf6a549dbd129f

Observation de0b357e-c288-41aa-af0f-e59281d439fe · outbound

This paper cites HuggingFace's Transformers: State-of-the-art Natural Language Processing.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain HuggingFace's Transformers: State-of-the-art Natural Language Processing

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T15:33:55.546557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:33:55.546557Z digest=sha256:1cfdf01fedfc3308b8ce4fb87af8499b7e9dda9d66c6785a66ba0761807fb946

Observation 8b16b78e-9cf1-419d-94fc-e7029f1716da · outbound

This paper cites Principled Reinforcement Learning with Human Feedback from Pairwise or $K$-wise Comparisons.

FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain Principled Reinforcement Learning with Human Feedback from Pairwise or $K$-wise Comparisons

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T15:33:55.702118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:33:55.702118Z digest=sha256:4cd19547fcfdf5213ca09b6502e11abbace7d29b5d2c1b4788308d40befc5b85

Pith citing papers

Observation 6834289c-1c52-4239-8fce-1a850878b6aa · inbound

AutoRubric-T2I: Robust Rule-Based Reward Model for Text-to-Image Alignment cites this paper.

AutoRubric-T2I: Robust Rule-Based Reward Model for Text-to-Image Alignment FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-20T12:13:16.128122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-20T12:11:23.775843Z digest=sha256:b13af3aeaf25e831ae0c5ca66c30fb2a7c6dc6946ec0bea30e6048cbed601b79

Observation db9b8703-7e0e-4549-9708-3d3e6b5dc41b · inbound

AutoRubric-T2I: Robust Rule-Based Reward Model for Text-to-Image Alignment cites this paper.

AutoRubric-T2I: Robust Rule-Based Reward Model for Text-to-Image Alignment FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-22T09:21:21.457051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T09:19:39.848194Z digest=sha256:3117fd25aedd8bf4e046537d9d7a3acebcb85d87b58129a547db954e83f2eb49