Pith. sign in

Paper Citation Record · LEDGER

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers

As of 10 August 2026, this Paper Citation Record lists 56 of 56 outbound references and 1 inbound Pith citation observation for arXiv:2502.03793.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.03793 v2

Coverage vector

measured 56 of 56 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T00:49:12.272682Z

measured 57 of 57 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:48:07.106359Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T17:48:07.604735Z

Reference resolution

56 of 56 outbound references displayed

  • verified exact3
  • verified fuzzy13
  • unresolved38
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation cc187f64-08d4-4506-8f5c-9dfb66ab5f66 · outbound

This paper cites Devlin, M.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Devlin, M

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.089123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.089123Z digest=sha256:3814dfee7b1d09e8e239fb60988acb66288a9723ac3f73ebe2325ee7968ac86b

Observation 22cdd5a5-e432-466c-adbc-cbac4b4dd012 · outbound

This paper cites Vaswani, N.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Vaswani, N

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:49:12.777086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T00:49:12.092704Z digest=sha256:186af8b5dfedc2447318a44dea81195f446f38f2b88e20ea977802750d256316

Observation 49986245-a2c9-4247-bb85-3dec3673329e · outbound

This paper cites Tulu 3: Pushing Frontiers in Open Language Model Post-Training.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Tulu 3: Pushing Frontiers in Open Language Model Post-Training

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.095682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.095682Z digest=sha256:41c4d8592b7da6db4e1cd32e4f1a7b49cea52a5259f22c9f925fb913b2384a99

Observation 400316d9-f446-46eb-a295-5d494dcfdad4 · outbound

This paper cites The Llama 3 Herd of Models.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers The Llama 3 Herd of Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.098714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.098714Z digest=sha256:e91af8dce579e21f727ba87f6a5ec784937646d192f84dab6ef5857bdafd21b3

Observation 55f7d166-2e43-4751-9fcb-89d6573861f8 · outbound

This paper cites Qwen2.5 Technical Report.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Qwen2.5 Technical Report

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.101823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.101823Z digest=sha256:e77bcac45e165b07baf9f69f48e84f136967d44b3fbef9a1c6809022039441fe

Observation ea9ebb7e-5310-4cad-ba3d-7e56cfcf2934 · outbound

This paper cites an unresolved cited work.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:49:12.770286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T00:49:12.105079Z digest=sha256:4ca520d5836b104fea866e00337fc301e1d6d9c01e009b0a3611e1509e7baa7d

Observation 04652211-c8e1-4e13-96eb-02bfcbb15305 · outbound

This paper cites Radford, J.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Radford, J

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:49:12.763255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T00:49:12.108155Z digest=sha256:e310fdb77fec023f5c562fb6de260a551f0f1c4cdb6f71fbe2572e27569aacad

Observation 00a61132-d8b6-42e9-b6f6-2ceac0fed9f1 · outbound

This paper cites Finetuned Language Models Are Zero-Shot Learners.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Finetuned Language Models Are Zero-Shot Learners

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.110889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.110889Z digest=sha256:4b630dff5076a8fbe6760c2a6e08db5a45fb836c0dcafaab7d91143f4485f407

Observation d2d252ac-0ae0-4771-a493-3480d848cc2a · outbound

This paper cites an unresolved cited work.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:49:12.755261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T00:49:12.117045Z digest=sha256:255dc9147dc9e44e3037297d59b5f1eec359df3f515eae72e9b71d70004930aa

Observation 384a8075-f251-4be6-8579-106cb6463d61 · outbound

This paper cites Warner, A.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Warner, A

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:49:12.746731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T00:49:12.120107Z digest=sha256:ad27664b8e91022c2fec73fd0d31b8b694af847dd6a21fb4ccb94058e04f78be

Observation 029cfc9e-e971-4cb0-afbc-c1fee4e52d77 · outbound

This paper cites Chain-of-Thought Prompting Elicits Reasoning in Large Language Models.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Chain-of-Thought Prompting Elicits Reasoning in Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.124370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.124370Z digest=sha256:a3db3b24b728e4da58504b4809a50434c384c31cdca761ebd0cc81c8a7ad38a3

Observation 42474086-e085-4513-bc04-7a813f4f7b3e · outbound

This paper cites an unresolved cited work.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.127371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.127371Z digest=sha256:5ca4970216ee4d48d3569ce457c8e6f0fad2ac9c6482a0788e62eb916dee2854

Observation 4f511a31-509a-484e-a37b-f6f478050150 · outbound

This paper cites Williams, N.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Williams, N

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:49:12.738655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T00:49:12.130541Z digest=sha256:5502243e4c159d0d68045703ec45052792e00885f9fca51a78a6a92e4d53dd1e

Observation 57a9767d-54d3-4d57-91c5-e4d6643aa27f · outbound

This paper cites Lewis, Y.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Lewis, Y

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.133110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.133110Z digest=sha256:5478de4478b458c7d9cc97f19dc501a2e2509c75fa3891d772f53519c1190b47

Observation 6c0f14fa-70c8-44af-b1f4-7ee5c8641108 · outbound

This paper cites an unresolved cited work.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:49:12.730589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T00:49:12.135882Z digest=sha256:4df4d6a703466b45f13bf1062a59562a57fc7ff6e233cad5fe649040de57c63e

Observation 81724653-4351-41d6-96b8-b1da6a36bc38 · outbound

This paper cites an unresolved cited work.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Unresolved cited work

Reference 17

Resolution
malformed identifier
no resolver link, observed 2026-08-09T00:49:12.138705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.138705Z digest=sha256:c50d2697cb971a7cf805a61e66ce7be23befc50ea0985e0b80629b3f1071f330

Observation 677a9b5e-a461-4546-8cea-3d2c65f71208 · outbound

This paper cites cloze procedure.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers cloze procedure

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:49:12.722993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T00:49:12.142256Z digest=sha256:77b4dfa38ff8ed22f5ba7600e44e2919969025cf17251bf6cba9c30b12762a4e

Observation 1883359b-0aa2-4ad8-b81f-60b4cc9357a9 · outbound

This paper cites Schick, H.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Schick, H

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.147694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.147694Z digest=sha256:960fa572553f720b0318e4444e12f2b279fe5450c4ad150a87bfe05bbe39dfd6

Observation 727558e6-826b-48f3-b14e-2ef42269f5aa · outbound

This paper cites BERTs are Generative In-Context Learners.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers BERTs are Generative In-Context Learners

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.150658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.150658Z digest=sha256:7d2221cadd489d3536e0b8b2555e9fa98d927e6ee5e7d051e0a2bb3f668c2a3b

Observation 5eb73e16-027b-4e3e-ab2d-7b9b91c7bc34 · outbound

This paper cites an unresolved cited work.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Unresolved cited work

Reference 22

Resolution
verified exact
doi, observed 2026-08-09T00:49:12.333868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T00:49:12.153789Z digest=sha256:2d9b87bf0acd96541e8c6ec83990f7cd8e87f30c5dca54475c0b6d8ef11d026e

Observation 9d14c827-305c-4ec8-827a-3e4a1f7c6089 · outbound

This paper cites DeepSeek-V3 Technical Report.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers DeepSeek-V3 Technical Report

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.156833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.156833Z digest=sha256:79abaf202f1c55b5967ea96f706eb5e77b2e3e33a8dfb0c7a8237a5562533e3a

Observation 2cc7376d-b3f4-4990-942b-f6a6e2989072 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.159808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.159808Z digest=sha256:3746e8ebaad449066f06c3c4d99732c71e12cda8b40cdcd29d5f3fa74e5dbc5f

Observation 49fb940e-5250-4f1d-a5cb-e951956d090b · outbound

This paper cites Qwen2 Technical Report.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Qwen2 Technical Report

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.162737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.162737Z digest=sha256:ae901c66d0f56dfd6c7316d18230f3be1fa587fae0dd08c6ad42f5f945f5c388

Observation 72e65d2c-17a6-4faf-a9f1-c353d719ebab · outbound

This paper cites Dolma: an Open Corpus of Three Trillion Tokens for Language Model Pretraining Research.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Dolma: an Open Corpus of Three Trillion Tokens for Language Model Pretraining Research

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.166031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.166031Z digest=sha256:b5168c65d0fc98b8c23ab647c50b1016a7e89722b5125c2af874ffc0609e4784

Observation e8469c3f-7a12-4714-ba72-5335a943d876 · outbound

This paper cites DataComp-LM: In search of the next generation of training sets for language models.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers DataComp-LM: In search of the next generation of training sets for language models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.169403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.169403Z digest=sha256:60b3747d47a788508d1010563d93f928eafa654c04f9cc7ab22a92afa306e08e

Observation 48009407-6cc4-4b6b-84fa-d505198f9747 · outbound

This paper cites The FineWeb Datasets: Decanting the Web for the Finest Text Data at Scale.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers The FineWeb Datasets: Decanting the Web for the Finest Text Data at Scale

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.172991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.172991Z digest=sha256:e59234b43d9d1a863403c2dea4cd3b8be0011d3d3c93bb232623ede201a22f5a

Observation b9053eac-b2a3-4d70-a394-df095850ad93 · outbound

This paper cites Smarter, Better, Faster, Longer: A Modern Bidirectional Encoder for Fast, Memory Efficient, and Long Context Finetuning and Inference.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Smarter, Better, Faster, Longer: A Modern Bidirectional Encoder for Fast, Memory Efficient, and Long Context Finetuning and Inference

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.176185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.176185Z digest=sha256:73b84f801dc77428c796b48bb230cfdcf28d11597f63de7a8375d15b3edc8dbe

Observation 9c65427e-5506-45f0-a57f-493555cb3691 · outbound

This paper cites an unresolved cited work.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:49:12.714808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T00:49:12.179707Z digest=sha256:ccdc9827c73dfe3faed51a53e5cb56b788819fce8fef562a6ef7798da9774c59

Observation fd3c91fe-112f-4c6c-aae9-2a9de888dd46 · outbound

This paper cites Zhang, Y.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Zhang, Y

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:49:12.706544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T00:49:12.183270Z digest=sha256:d813728af4e84c892ff1c7a524f6f06d90c2c1a6bf10a65e78ea19d9d291f2f9

Observation 1575c2c5-855c-415a-8bec-64d3902e0bae · outbound

This paper cites Javaheripi, S.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Javaheripi, S

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:49:12.698090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T00:49:12.186606Z digest=sha256:5d16c5ce0bb17adfe57d96fc1da935b252d54952f96c50f37128c82a6fa09c05

Observation b130bb73-1fce-494b-aaf2-bf70cd1024e0 · outbound

This paper cites Mamba: Linear-Time Sequence Modeling with Selective State Spaces.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Mamba: Linear-Time Sequence Modeling with Selective State Spaces

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.189956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.189956Z digest=sha256:6f590055d2af2e5734c2fea6c7848bea26152650e17ac69140ec2430af98c1c1

Observation 49d80b63-c591-4c7c-978a-b3b79a642df8 · outbound

This paper cites an unresolved cited work.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:49:12.690578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T00:49:12.193102Z digest=sha256:62283afd19f26bcc7dddc3b53fb27802a588e3fb0ff90edfc225ddba1e3be0ad

Observation c08affc9-d785-4f9a-968a-9606dcf46c3c · outbound

This paper cites DeBERTaV3: Improving DeBERTa using ELECTRA-Style Pre-Training with Gradient-Disentangled Embedding Sharing.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers DeBERTaV3: Improving DeBERTa using ELECTRA-Style Pre-Training with Gradient-Disentangled Embedding Sharing

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.195565Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.195565Z digest=sha256:a7583556e2d5cf15bc526c9a6532a279493bf992d514fba4efbf81d907bc2dc9

Observation 67a90a28-3383-4623-b097-725b80457f00 · outbound

This paper cites Sutskever, O.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Sutskever, O

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:49:12.683142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T00:49:12.198723Z digest=sha256:fededaf6826ced30a9605a69bfa645db964deb79ce162dc733b98e88edd83259

Observation 9c4216a1-3d61-45b9-a617-0cdd483ee7f9 · outbound

This paper cites Ra ffel, N.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Ra ffel, N

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.201569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.201569Z digest=sha256:f8464aa8b5d130263a3c2acd130d9cdd4564aede22352a846632e8080506c637

Observation eabbf38e-33b3-4748-91fc-a3fd99de6f55 · outbound

This paper cites an unresolved cited work.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Unresolved cited work

Reference 38

Resolution
verified exact
doi, observed 2026-08-09T00:49:12.312478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T00:49:12.204286Z digest=sha256:aafeaaf7103c655e20970933f0141b5e10d4c5091d7317d5da7d5b087c55464c

Observation 7c060403-72a8-45f4-97ea-5dd4431e0d1f · outbound

This paper cites Longpre, L.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Longpre, L

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:49:12.670374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T00:49:12.207043Z digest=sha256:f948d2a08d2969b929547560fd394d982cf7400b9b8c18658023d05dbb7f798c

Observation d1d0f590-36a1-4de5-ae50-26a263b8076d · outbound

This paper cites Hendrycks, C.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Hendrycks, C

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:49:12.662642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T00:49:12.210834Z digest=sha256:482af237afbafe4ebf45c199f6eaa346b1df2fdd9287e0deff19821da848f8bd

Observation cecd3773-fd12-44e2-90f6-cbc869d0af81 · outbound

This paper cites Challenging BIG-Bench Tasks and Whether Chain-of-Thought Can Solve Them.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Challenging BIG-Bench Tasks and Whether Chain-of-Thought Can Solve Them

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.213813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.213813Z digest=sha256:8704431e3fcc73bbea079d41bb8327032503c3bbe0a4a0866241e0c5d106f8d1

Observation 6fef0aae-67cd-4a5a-8bf2-3240be7a6197 · outbound

This paper cites Srivastava, G.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Srivastava, G

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.217095Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.217095Z digest=sha256:811c56693e0b76839bfd8d1dc31b2e1ce581dd678e3e24ae3cffaa54728b691a

Observation 70ce846d-d0da-4fce-9eec-c1496881593b · outbound

This paper cites Khattab, M.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Khattab, M

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:49:12.650911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T00:49:12.220142Z digest=sha256:49c66bb303f13ff4babdfca908828f7e496ce04df326e5ee76c451bc6db4f78e

Observation 4b016110-2dab-4397-bf2e-92efff9fd97a · outbound

This paper cites Text Embeddings by Weakly-Supervised Contrastive Pre-training.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Text Embeddings by Weakly-Supervised Contrastive Pre-training

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.223389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.223389Z digest=sha256:7496c6e85bb1af2ebff0b42c78e0e286f54a4a6a5d75be5a9f4311ae39c4a976

Observation 31aa18ef-6a75-4636-9511-47909dc375cb · outbound

This paper cites Zhang, et al., Jasper and stella: distillation of sota embedding models, arXiv e-prints (2024) arXiv–2412.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Zhang, et al., Jasper and stella: distillation of sota embedding models, arXiv e-prints (2024) arXiv–2412

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:49:12.641399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T00:49:12.227145Z digest=sha256:77aba9d16946402c2125b7a8ca320f67243e05e04bcbbb5bf4da36e3ca177e0b

Observation 56b514b6-0342-4bc5-ac0d-acb81f14916a · outbound

This paper cites MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.234415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.234415Z digest=sha256:31b00415adf5d6629b01038e14f9c5f4ec3c1c4d0ce347a40c1f5f68590a787c

Observation ff5c76f0-f918-4ccc-93fa-93725ff048de · outbound

This paper cites RAFT: A Real-World Few-Shot Text Classification Benchmark.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers RAFT: A Real-World Few-Shot Text Classification Benchmark

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.237357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.237357Z digest=sha256:40b4a4b5f9121a19136bf525e59cab9184894f88eaad2b551b4aeae70e522557

Observation c26a35cf-1d59-43d3-aedd-873f115ce4ca · outbound

This paper cites Gurulingappa, A.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Gurulingappa, A

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.240445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.240445Z digest=sha256:6b02c21094784cc0ef2393da896e4320dbe6b82509b8225ec3cbde91cf0503c0

Observation 30f134f6-211e-41aa-bc9c-888d74ff69e0 · outbound

This paper cites Vajjala, I.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Vajjala, I

Reference 49

Resolution
verified exact
doi, observed 2026-08-09T00:49:12.300315Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T00:49:12.243751Z digest=sha256:5ae0370fac7f0ba207cfb46b50276f51e8c45e60e585b2fe9d6c973771a80952

Observation 2bf2b1d9-718d-44bf-b78a-68021688fab4 · outbound

This paper cites an unresolved cited work.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Unresolved cited work

Reference 50

Resolution
malformed identifier
no resolver link, observed 2026-08-09T00:49:12.246217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.246217Z digest=sha256:bf8c9df4db510920f38001995393bb180b1c852434d107b7260b544e1557ce62

Observation d012e6c7-8181-4285-ae60-d561652ec04f · outbound

This paper cites Zhang, J.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Zhang, J

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:49:12.633224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T00:49:12.248745Z digest=sha256:676b4f428fda2a1384a5bca154119aa465a686b76f2a50a311734c785ac7c7f4

Observation f94c6c45-71ce-4498-bce1-cc7bfe0b2175 · outbound

This paper cites an unresolved cited work.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:49:12.625095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T00:49:12.252185Z digest=sha256:f119d85038d3d653040a913aa78a338f85da45c5b1c46d2d58ac98b1c364c822

Observation 853fa846-d311-4f34-87a9-536cb65c7633 · outbound

This paper cites RoBERTa: A Robustly Optimized BERT Pretraining Approach.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers RoBERTa: A Robustly Optimized BERT Pretraining Approach

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.255533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.255533Z digest=sha256:accb2945fed5c05e0c8521e7330f2aef0ab3be25339efcc7c6197eb9bc3e5b85

Observation 37139d8d-fe31-4959-8b98-5aefdd56cb93 · outbound

This paper cites an unresolved cited work.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Unresolved cited work

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.259260Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.259260Z digest=sha256:581299466dc0bdc3a0d50fd518c54a3d7c12aa2529a9555d37498675be3f9c53

Observation eddd7f55-8fb1-4944-80ca-562e3ffc17f2 · outbound

This paper cites an unresolved cited work.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:49:12.616581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T00:49:12.262838Z digest=sha256:6df3574282ece584abc4d02fbdf2875816a9ba13a61152c18efa7b0f153237c5

Observation 02a1e252-c7d2-49dd-9dbb-d72c34d7d4b6 · outbound

This paper cites Hermes 3 Technical Report.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Hermes 3 Technical Report

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.266170Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.266170Z digest=sha256:463e940971f1b00222b3d68e421b6827773c52b22ea9b9fa9a9f71430911b07f

Observation 0a211ce9-50bb-4526-8191-9360ed615c75 · outbound

This paper cites A Survey on In-context Learning.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers A Survey on In-context Learning

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.269561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.269561Z digest=sha256:9749950d685c51d1225b05494c95bbe93b77f466609f87bc2201a0399201603d

Observation 75fdcb7c-54bd-4507-95c5-3b735ac1ead5 · outbound

This paper cites an unresolved cited work.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:49:12.609062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T00:49:12.272682Z digest=sha256:109ac01582fea538985454aca23077396859da4e24f6663dc1fdb8149d713b00

Pith citing papers

Observation f86633c8-6b4d-4265-b4bc-f7dbf11bcbe6 · inbound

Tiny Reward Models cites this paper.

Tiny Reward Models It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-08-06T17:48:07.607922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T17:48:07.106359Z digest=sha256:f256a749a3a8e2e8af9ea78e5814a5e61f89917abdd6bbea155e286bec874b60