Pith. sign in

Paper Citation Record · LEDGER

Optimising Language Models for Downstream Tasks: A Post-Training Perspective

As of 19 August 2026, this Paper Citation Record lists 100 of 277 outbound references and 0 inbound Pith citation observations for arXiv:2506.20917.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.20917 v1

Coverage vector

measured 100 of 277 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T22:44:44.107619Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 277 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved98
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4b73827d-5f4f-4f31-819b-a794dfb137a7 · outbound

This paper cites Nemotron-4 340B Technical Report.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Nemotron-4 340B Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:42.951443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:42.951443Z digest=sha256:6b14611ea996789aad7a0c4c0b401f8000772f4327a0cc0c7f750967176e147d

Observation 8d58da42-ab95-45b2-a929-d5c97fcc36c8 · outbound

This paper cites Publicly available clinical BERT embeddings.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Publicly available clinical BERT embeddings

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:42.966717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:42.966717Z digest=sha256:dc1712e7b3e36fe4c82dcff0b895c368378c28df83d0faec8719bdcadef01c77

Observation d875b45e-72aa-4a79-97fe-5717c8862855 · outbound

This paper cites Reid, Stephen Gould, and Anton van den Hengel.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Reid, Stephen Gould, and Anton van den Hengel

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:42.977418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:42.977418Z digest=sha256:8688181ab749674a5a43507edfb4c215e03efb6a96e60f40cecf36687b36304d

Observation de2b65ba-e9fa-490f-b0aa-9524bb270450 · outbound

This paper cites Pseudo-Labeling and Confirmation Bias in Deep Semi-Supervised Learning.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Pseudo-Labeling and Confirmation Bias in Deep Semi-Supervised Learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.098284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.098284Z digest=sha256:ca78bc2f4e2c27276a780c2fd8c9f6e0537a75326aa9034e83edf67d12441ca8

Observation 13c2475a-6020-4db0-bf03-fbd2d68919fc · outbound

This paper cites Tran, Dara Bahri, Jianmo Ni, Jai Prakash Gupta, Kai Hui, Sebastian Ruder, and Donald Metzler.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Tran, Dara Bahri, Jianmo Ni, Jai Prakash Gupta, Kai Hui, Sebastian Ruder, and Donald Metzler

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.199897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.199897Z digest=sha256:9f3f5000a093113fa68ededdd593a50bbc42f76e47a884f72e88b4fa40eae8fa

Observation b706fb71-2eaa-44fe-b1ee-ab20c66b1dcf · outbound

This paper cites A robust self-learning method for fully unsupervised cross-lingual mappings of word embeddings.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective A robust self-learning method for fully unsupervised cross-lingual mappings of word embeddings

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.209437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.209437Z digest=sha256:90a55ec9e8c072f1cb13d8cbdd5975075f4dbb8ba279c1898941b387dfa04669

Observation 7c41c763-bf8e-4833-8255-181336720d8c · outbound

This paper cites ATTEMPT: Parameter-efficient multi-task tuning via attentional mixtures of soft prompts.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective ATTEMPT: Parameter-efficient multi-task tuning via attentional mixtures of soft prompts

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.213678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.213678Z digest=sha256:4c239b8cdedbfe863b570f2d8c204ddfeece5a26a9057981b26853d3dbcd51a1

Observation cb805e14-e2aa-457e-93b7-6ac967060ede · outbound

This paper cites Program Synthesis with Large Language Models.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Program Synthesis with Large Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.218178Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.218178Z digest=sha256:58dccd2917285742966dbe5407571df929272842d2bb5ef6b790f7af4de4d1b1

Observation 326ed919-0310-4d12-96cc-abb66d392730 · outbound

This paper cites A general theoret- ical paradigm to understand learning from human preferences.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective A general theoret- ical paradigm to understand learning from human preferences

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.232106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.232106Z digest=sha256:e988d8ae5f1124c3a06e982f708214016b2de55a247717d6562fa0ac8af56ed2

Observation c588d5d2-b57b-4974-9429-dfde47e4da1b · outbound

This paper cites Qwen Technical Report.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Qwen Technical Report

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.277409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.277409Z digest=sha256:c0cd2a2110e40b8009cad896bcfbe6ae7a8e50ca81383504bae729fd5e401d2f

Observation 73ab6ab2-a9fc-4a1f-9b19-3e278d9fb56f · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.346100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.346100Z digest=sha256:f529e157f46f1cb3ce9ee62edfae971310c3f65f38fbc65d2b89cea0f759b385

Observation 1d63966b-582c-458a-a353-c8ce30255fa7 · outbound

This paper cites Brain power.Proceedings of the National Academy of Sciences, 118(32):e2107022118, 2021.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Brain power.Proceedings of the National Academy of Sciences, 118(32):e2107022118, 2021

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.421930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.421930Z digest=sha256:87f28490afcc58ca1c80698805e8ee7ad1075245f0edda931f892286fa070451

Observation c1dcd433-9c1f-4cc8-9b5a-81a82ff68d29 · outbound

This paper cites SciBERT: A pretrained language model for scientific text.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective SciBERT: A pretrained language model for scientific text

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.483528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.483528Z digest=sha256:6882c110495cbf928bc5ac887ae1fa3b4a647edc5bef02eb9a5c6bcc12f14e83

Observation 45cd9730-f44b-49f8-b7a3-f76a57fa7777 · outbound

This paper cites BitFit: Sim- ple parameter-efficient fine-tuning for transformer-based masked language- models.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective BitFit: Sim- ple parameter-efficient fine-tuning for transformer-based masked language- models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.535208Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.535208Z digest=sha256:dc8fd18dc42ab343c286c58ff99c011fb9ca403ee04fd9271ebd53a6f3928356

Observation f4135c37-405d-45af-8bad-6890dd4292b4 · outbound

This paper cites Bender, Timnit Gebru, Angelina McMillan-Major, and Shmargaret Shmitchell.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Bender, Timnit Gebru, Angelina McMillan-Major, and Shmargaret Shmitchell

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.597280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.597280Z digest=sha256:34d4b8d7933a4593e801a7a53c0dea5186367e6612fac3d07268c20de70964fb

Observation fe5548a5-b899-4d00-bc42-83efd4540425 · outbound

This paper cites Bender and Alexander Koller.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Bender and Alexander Koller

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.732143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.732143Z digest=sha256:40f406b53a49195e123c9ab06a2b7a42c3112f3cf378b8ac1fc84f852b717c15

Observation e892c41f-67e4-4c8e-b2f6-0e9409c15c00 · outbound

This paper cites Curriculum learning.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Curriculum learning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.735683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.735683Z digest=sha256:f784b13443e3558ff9782f9e0f8482a99cc94c66caa49fb1cef905fb35581c97

Observation 8ba42f1b-4a05-44d1-a5f5-188fe7c5356d · outbound

This paper cites The fifth PASCAL recognizing textual entailment challenge.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective The fifth PASCAL recognizing textual entailment challenge

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.738874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.738874Z digest=sha256:9a96204c20756bd9e8447352ec58e3e0fe694a8f87542533dd43c708462cd8e8

Observation a0832af2-56e1-47af-8bba-2b9b04704280 · outbound

This paper cites Cubuk, Alex Kurakin, Kihyuk Sohn, Han Zhang, and Colin Raffel.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Cubuk, Alex Kurakin, Kihyuk Sohn, Han Zhang, and Colin Raffel

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.742786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.742786Z digest=sha256:ebfac967c9c6e720b87eb0421a7551240b095353402fdf2e3a17712c3ca16533

Observation 3c194945-8026-4ce1-ab53-234328ee7e19 · outbound

This paper cites Goodfellow, Nicolas Papernot, Avital Oliver, and Colin Raffel.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Goodfellow, Nicolas Papernot, Avital Oliver, and Colin Raffel

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.746891Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.746891Z digest=sha256:678369de9d60012c7cebedecd95f55cb481edb87f44b93e1f1da08440c0424de

Observation 49f58851-1f47-4f28-8cd0-f06bb0480165 · outbound

This paper cites Adamatch: A unified approach to semi-supervised learning and domain adaptation.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Adamatch: A unified approach to semi-supervised learning and domain adaptation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.750663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.750663Z digest=sha256:075323228a1d61671540db2a90ef75dfcfeb3965d705fde57600eef368e9a46a

Observation 647b04ba-67f9-4605-8227-34babddd1e3c · outbound

This paper cites Shih, Yejin Choi, and Daniel Marcu.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Shih, Yejin Choi, and Daniel Marcu

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.754753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.754753Z digest=sha256:f95b5e9c07765af0a234c97ef8e4cf2e7b1b1dc4933a1758b80b8de7ed354806

Observation 7695e2b0-2af6-4ec2-a284-36335d8dfe45 · outbound

This paper cites PIQA: reasoning about physical commonsense in natural language.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective PIQA: reasoning about physical commonsense in natural language

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.759066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.759066Z digest=sha256:6b86724eac8b40c42dc1afa52dd8a78a8bb546f17ab4c30144b29dbb073e72e7

Observation 1afe4fb7-f00c-4da2-bf75-566cd5d388d2 · outbound

This paper cites Bowman, Gabor Angeli, Christopher Potts, and Christopher D.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Bowman, Gabor Angeli, Christopher Potts, and Christopher D

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.763474Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.763474Z digest=sha256:746397aac869ed9e2ffaa68a854471e0f81b1e102c719d81471b16879183c2a2

Observation 88dc5620-7baf-42b8-ba46-bec13b826d0b · outbound

This paper cites an unresolved cited work.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.768026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.768026Z digest=sha256:680ddfe61fb8f772b457e3d1432198c3c77d498842992187d465ef28707b17fd

Observation 0ba7bc8b-8f5b-4ba0-aee8-fd2cf466cfaf · outbound

This paper cites Semi-supervised semantic role labeling with cross-view training.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Semi-supervised semantic role labeling with cross-view training

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.772139Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.772139Z digest=sha256:98811371ad85c1503c9178585ff0667f29e051750e9d0969ebf965ff82bc88c5

Observation e9b5ec18-87be-4a50-aaf6-7afa6c962e28 · outbound

This paper cites Findings of the 2009 Workshop on Statistical Machine Translation.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Findings of the 2009 Workshop on Statistical Machine Translation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.777079Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.777079Z digest=sha256:e729cfd5ab9fb799f207adf934feb317b50038ebec407c0a62183151199834e4

Observation 64557a6d-0750-4240-86e8-c773f78a6539 · outbound

This paper cites Extracting training data from large language models.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Extracting training data from large language models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.781782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.781782Z digest=sha256:06f77363e867303e1e50f5c810ed556762bc40d6689ee635299c12e827269f7f

Observation 1dd787d0-4e16-421b-b31d-acd77c5e2bf0 · outbound

This paper cites SemEval-2017 task 1: Semantic textual similarity multilingual and crosslingual focused evaluation.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective SemEval-2017 task 1: Semantic textual similarity multilingual and crosslingual focused evaluation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.790303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.790303Z digest=sha256:0265452b76a8318d0150ebd7b93258b039a7dcaca56ff6be221ba6ce16a31d69

Observation f8595921-d963-4d70-9f72-775a66f409b4 · outbound

This paper cites Importance of semantic representation: dataless classification.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Importance of semantic representation: dataless classification

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.794286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.794286Z digest=sha256:813bd2f29ebb70dcf08d764d8f259cb68ccda41127b198ee4829ff26bf79f295

Observation afb4a726-e2b8-409d-8cf9-fafd94f1b3db · outbound

This paper cites Semi-supervised BIBLIOGRAPHY 121 learning (chapelle, o.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Semi-supervised BIBLIOGRAPHY 121 learning (chapelle, o

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.798701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.798701Z digest=sha256:3d9fece55c6eba595476f58bcce913a36244c6b16e1504d47dffb058c3917c54

Observation a6fe9601-807c-43b7-8236-ab98669c0721 · outbound

This paper cites Code alpaca: An instruction-following llama model for code generation.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Code alpaca: An instruction-following llama model for code generation

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.802803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.802803Z digest=sha256:ab08765de60894922e620e72c6487cdafcdab5959e7e503366c1bb79fb4330ed

Observation 514b73ed-c09b-4d89-b55f-282201ca692b · outbound

This paper cites Debiased self-training for semi-supervised learning.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Debiased self-training for semi-supervised learning

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.806846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.806846Z digest=sha256:d1f8a22a794071c33eb677e7b168479879477171a05e7269465d8d4f62782ff8

Observation 98443fdc-0446-422f-b965-4ec3a5b44748 · outbound

This paper cites Unseen filler generalization in attention-based natural language reasoning models.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Unseen filler generalization in attention-based natural language reasoning models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.810858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.810858Z digest=sha256:94c6cedaeae2fc6c0aefc51fa1e095c242a28e8420931840a48cb9b8d774e3bf

Observation 386a0536-a5ec-470d-b323-6eda5ce5e618 · outbound

This paper cites TOUCHDOWN: natural language navigation and spatial reasoning in visual street environments.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective TOUCHDOWN: natural language navigation and spatial reasoning in visual street environments

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.814900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.814900Z digest=sha256:2c78caf1f08c638bf7aee21813720e6e2f6a286a34df112597f37ec0ef813049

Observation b962e938-27ce-4e8f-aa7c-9faf82ffd4e7 · outbound

This paper cites MixText: Linguistically-informed interpolation of hidden space for semi-supervised text classification.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective MixText: Linguistically-informed interpolation of hidden space for semi-supervised text classification

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.818998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.818998Z digest=sha256:126c6cff2b46e73737e72e21b7fa4906dec6d1dc8bbf0852a33473857737bc72

Observation f9e4a0c8-d035-4a5e-88cc-d8cf8c5af729 · outbound

This paper cites Alpagasus: Training a better alpaca model with fewer data.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Alpagasus: Training a better alpaca model with fewer data

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.823064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.823064Z digest=sha256:cb6f8748bf54e6ca951d4be43a3915553ce6202bdadd43bd50d36a1b05fbc69b

Observation 9c8fff1d-b1a5-4387-87e1-cd0d4ad17d28 · outbound

This paper cites Evaluating Large Language Models Trained on Code.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Evaluating Large Language Models Trained on Code

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.827536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.827536Z digest=sha256:044c57e42bc225fe6e5f7b414a7fdde5546714060c4e47dd05aa5c6e9e69e523

Observation 7f06129c-3df5-41a6-b17c-859f4f8c4064 · outbound

This paper cites Microsoft COCO Captions: Data Collection and Evaluation Server.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Microsoft COCO Captions: Data Collection and Evaluation Server

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.832810Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.832810Z digest=sha256:88eced2552a0cfa9de3dc4eb772d1af4dc15337f77545cddf94460cfb1077dcc

Observation 4ceffca8-0791-4934-b47d-4c6d2435dd28 · outbound

This paper cites Adapt- ing language models to compress contexts.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Adapt- ing language models to compress contexts

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.837349Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.837349Z digest=sha256:14f87744c33de7a80f950b85de20f1caa6444420841b2f20e1a2a7a4c3892ac6

Observation fa025a4d-3938-4dc7-ae30-02f10aee1c4f · outbound

This paper cites Gonza- lez, Ion Stoica, and Eric P.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Gonza- lez, Ion Stoica, and Eric P

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.845371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.845371Z digest=sha256:389640957c7aea59c6fcfd13533d2b95915716ba1da7e552fc4ee2c2b05c1701

Observation 9c40c8b1-87b4-48d4-af4d-f8c49374f510 · outbound

This paper cites PaLM: Scaling Language Modeling with Pathways.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective PaLM: Scaling Language Modeling with Pathways

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.849926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.849926Z digest=sha256:29ef5371b71d75b2756dd1f2db9761c3b3ba50b4dad0d173646ee659f4a6d1e2

Observation 45174e3b-1722-46a6-a373-2527451532a5 · outbound

This paper cites Deep reinforcement learning from human preferences.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Deep reinforcement learning from human preferences

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.854454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.854454Z digest=sha256:14337fd7cd626a4599930362fe6d1dbd559f3463f35e957e17bcfd4e3f2d9f8b

Observation d17e3aa2-e9d1-4863-9736-8477ecfe52a5 · outbound

This paper cites Chi, Jeff Dean, Jacob Devlin, Adam Roberts, Denny Zhou, Quoc V.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Chi, Jeff Dean, Jacob Devlin, Adam Roberts, Denny Zhou, Quoc V

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.858540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.858540Z digest=sha256:45ab1250253fe097699517629150bedde869c0dd550c66642bfd67ba2da64a82

Observation 8acdc684-4fc2-4f76-b0dc-e2f9b1611de0 · outbound

This paper cites BoolQ: Exploring the surprising difficulty of natural yes/no questions.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective BoolQ: Exploring the surprising difficulty of natural yes/no questions

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.863023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.863023Z digest=sha256:bedcd1d2c533cfdc38a84a99a07933e5686c54d286eb53d745579df669b29299

Observation 6e83cee6-2d08-4b0f-ba24-1a8539d6dc57 · outbound

This paper cites Manning, and Quoc Le.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Manning, and Quoc Le

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.867863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.867863Z digest=sha256:2fd086c6bff7e22ee035a9c3aaf093381f3d3ddd77b067a688710b6f550ddbdd

Observation b4666c98-7120-4c9d-be79-d49331aab7b6 · outbound

This paper cites Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.871934Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.871934Z digest=sha256:678363ca675806167752c7b09233e6586e016ce585c2abed6a227be43b73888a

Observation 6cacfbce-4e24-404a-956c-d73ec416b0ad · outbound

This paper cites Training verifiers to solve math word problems, 2021.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Training verifiers to solve math word problems, 2021

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.876140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.876140Z digest=sha256:073d91d28ebf9166b776ef87c07aad59410e90b4bdbed456aef7219724c0e402

Observation 3dc26713-97e4-4cf2-bb77-6646961dcab3 · outbound

This paper cites Free dolly: Introducing the world’s first truly open instruction-tuned llm, 2023.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Free dolly: Introducing the world’s first truly open instruction-tuned llm, 2023

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.880797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.880797Z digest=sha256:147853467a858d3571f0c95b7ae8d4fff3e9ca01cd0456261f8e40fed139ae68

Observation 4cd63697-8582-43e4-b949-2fed3f54d4b6 · outbound

This paper cites The PASCAL recog- nising textual entailment challenge.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective The PASCAL recog- nising textual entailment challenge

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.884647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.884647Z digest=sha256:9719962540320f67c55288b74f80e7eeac83085c10b1a03c8a2c7cdc11dcb81c

Observation 03f88398-73a0-4d54-85f4-752cafe668ed · outbound

This paper cites Flashattention-2: Faster attention with better parallelism and work partitioning, 2023.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Flashattention-2: Faster attention with better parallelism and work partitioning, 2023

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.888406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.888406Z digest=sha256:db73b4d86d4a3c33803567eac503a13e8c38b1fc4b480bf190013ff5e3aee2cf

Observation 51c23cfc-d3dd-4812-8719-9751fa392809 · outbound

This paper cites The commitmentbank: Investigating projection in naturally occurring discourse.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective The commitmentbank: Investigating projection in naturally occurring discourse

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.892250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.892250Z digest=sha256:4fdc473cee5dce17293df74c7bfa0f5480341d552d15e3c11a43cec40a4b946b

Observation 8cd2fcdd-14b9-44cc-bb5c-ffb64dc62880 · outbound

This paper cites Universal transformers.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Universal transformers

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.896950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.896950Z digest=sha256:deb3bf3e631ed8c5f80a676d4d59972d3236dab542cc32a7c381f15d0a028654

Observation 78874513-6b27-49cc-b158-b38ccc5f947a · outbound

This paper cites BERT: Pre-training of deep bidirectional transformers for language understanding.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective BERT: Pre-training of deep bidirectional transformers for language understanding

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.905116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.905116Z digest=sha256:e013930138b8e45830363943b545c9d38abb50722a34d16d7710b5d28a3a2f56

Observation 0f4a975a-d3e1-4246-83c7-65ffeddc1ea9 · outbound

This paper cites Attention over learned object embeddings enables complex visual reasoning.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Attention over learned object embeddings enables complex visual reasoning

Reference 57

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T22:44:45.601668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:44:43.909542Z digest=sha256:fecf255ce20064c00f27908b628e279da258f8bd71fa63b62a10710d6ca88e25

Observation 485d0deb-e583-40a7-b3aa-d221baf29a60 · outbound

This paper cites Dolan and Chris Brockett.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Dolan and Chris Brockett

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.917917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.917917Z digest=sha256:47f0375e8926a2609837ca9a3f3dba4af23edfca1ac5b2f7c60480dba818a12d

Observation c54d89f2-42da-4a6f-bc85-75bd718958c9 · outbound

This paper cites A robust self-learning framework for cross- lingual text classification.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective A robust self-learning framework for cross- lingual text classification

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.921533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.921533Z digest=sha256:f754fa98d0f1eb4e916737a78c2385bab6e3c1abaf763b2cb9230ceeb9f657f1

Observation b1cf96da-04a9-4bfd-82c9-07bc2ef89990 · outbound

This paper cites The Llama 3 Herd of Models.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective The Llama 3 Herd of Models

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.925514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.925514Z digest=sha256:b480fc1c3e3000cdbb7b25471b071bfe1f042861d066b481b9506417efc2ef51

Observation 454af4f4-62f4-4af8-bcea-b3ab0bf330a0 · outbound

This paper cites SearchQA: A New Q&A Dataset Augmented with Context from a Search Engine.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective SearchQA: A New Q&A Dataset Augmented with Context from a Search Engine

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.929753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.929753Z digest=sha256:682933a7e089cf36fcef9128cd0a1670274a54e43d388cf827018348d736476c

Observation ca3c2e8d-9098-4958-a20d-fb04887dc98f · outbound

This paper cites MRQA 2019 shared task: Evaluating generalization in reading com- prehension.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective MRQA 2019 shared task: Evaluating generalization in reading com- prehension

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.934143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.934143Z digest=sha256:c416f9e9d60404481b32102b765cdd4022e045de90c4c0d6956c1594029a4f81

Observation 51d56377-32f5-4d7f-9ec8-b748e575d177 · outbound

This paper cites Making pre-trained language models better few-shot learners.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Making pre-trained language models better few-shot learners

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.938432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.938432Z digest=sha256:96a0b0a2c4f2c65a43b206315eb2f00ad6e0998ef264734ff0912cac18e0a2a6

Observation fd02e4cf-c60f-4c36-9207-7682117dccb2 · outbound

This paper cites Zero-shot text classification with self-training.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Zero-shot text classification with self-training

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.943172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.943172Z digest=sha256:73618ff5014dce7e89c0f9cfbd2860acb361377754e5889f404218adecb1756b

Observation f6a792ca-a239-401e-8b3e-020e0c206176 · outbound

This paper cites The third PASCAL recognizing textual entailment challenge.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective The third PASCAL recognizing textual entailment challenge

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.951585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.951585Z digest=sha256:f3b461359e568e72416b2b67309b06ae9c2fed159369364cd68f3a175a865856

Observation dc6b9d1b-277a-4708-9c8f-73a7cb4f0847 · outbound

This paper cites PARS: Pseudo-Label Aware Robust Sample Selection for Learning with Noisy Labels.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective PARS: Pseudo-Label Aware Robust Sample Selection for Learning with Noisy Labels

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.955463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.955463Z digest=sha256:ad2cd9c70d8485358ebc0fe378da2b06f36046344ac8cb3f3e1c94872329e343

Observation a6e7cd32-27b4-4356-92b5-47c15418a18c · outbound

This paper cites Making the V in VQA matter: Elevating the role of image under- standing in visual question answering.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Making the V in VQA matter: Elevating the role of image under- standing in visual question answering

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.960188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.960188Z digest=sha256:41203a965782560f1a323698b66cdda8ead984bf198acc63ce86a1b9423f1eb7

Observation 5f8850c1-6ad4-47e1-80d6-487256dba4e4 · outbound

This paper cites Semi-supervised learning by en- tropy minimization.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Semi-supervised learning by en- tropy minimization

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.964317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.964317Z digest=sha256:ca434db8271af2f2aed0fdbe7036590f0aae8d3cb26d2db3db60c0d6955499b8

Observation ef3006f6-2adf-47dd-bd74-2b0d2a27595b · outbound

This paper cites Projected language models: A large model pre-segmented into smaller ones.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Projected language models: A large model pre-segmented into smaller ones

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.968612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.968612Z digest=sha256:625cbaaf9e523b7a0fcb9a7ebf83594c337558556aff8a933fdaf8c079dc1374

Observation c06839e4-2587-4c38-9d9d-19c58db54669 · outbound

This paper cites PPT: Pre-trained prompt tuning for few-shot learning.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective PPT: Pre-trained prompt tuning for few-shot learning

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.973015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.973015Z digest=sha256:656bdfc5301f6e47faba4abe772b41fe20260fed708c35c9d12d8dfe15810fb2

Observation bde65ae0-cc71-4d47-93ad-2d76f9fad05d · outbound

This paper cites Textbooks Are All You Need.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Textbooks Are All You Need

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.977337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.977337Z digest=sha256:521fd91fbd4808e2b98c460da367dd37db0de57970dbce68fd19018e81591907

Observation 8535b9b1-5e36-4d82-a403-de8ddcb09be5 · outbound

This paper cites Parameter-efficient transfer learning with diff pruning.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Parameter-efficient transfer learning with diff pruning

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.982011Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.982011Z digest=sha256:a22f4f724b2035c66723d72b8f0904e4fe06371b30ab31f0b0226ed1b555f4a6

Observation 24285ebd-c1ae-4116-b986-77a679a384a9 · outbound

This paper cites an unresolved cited work.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Unresolved cited work

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.986757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.986757Z digest=sha256:7aa3519c827f8fda0a797a2d6c3605f8e11e39d1d31fdf7dc0293ac3d273d298

Observation 15de1586-1f21-4d33-b0fe-9ee0b4185d74 · outbound

This paper cites an unresolved cited work.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Unresolved cited work

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.990606Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.990606Z digest=sha256:f0ca58b4b2b9dd818f3fe943ae00338ae14c2c7e55bac5c6bc12a4962f81d05b

Observation f05badb1-233e-40c7-a630-7b8038610c34 · outbound

This paper cites W ARP: Word-level Adversarial ReProgramming.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective W ARP: Word-level Adversarial ReProgramming

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.994748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.994748Z digest=sha256:deb8d931ba059419beeb762053d7f9ee89ed06f83f3c93303121276337906f8e

Observation 8a42452f-7e20-4234-834c-93aef0b19e51 · outbound

This paper cites ToxiGen: A large-scale machine-generated dataset for adversarial and implicit hate speech detection.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective ToxiGen: A large-scale machine-generated dataset for adversarial and implicit hate speech detection

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:43.998737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:43.998737Z digest=sha256:628c1bb330a9d0b1abb817b3d200c34f870950b9c4476e078beede32f93cc4e6

Observation 508e78f9-fa73-48d0-9ba2-bb3920c27b89 · outbound

This paper cites Preserving Pre-trained Features Helps Calibrate Fine-tuned Language Models.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Preserving Pre-trained Features Helps Calibrate Fine-tuned Language Models

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:44.002754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:44.002754Z digest=sha256:9b404a46ce3049261bcf063333da6702131fc01608ea122f1f0128352adb4819

Observation 96c7696e-858c-4660-95cc-c48eda6dea5a · outbound

This paper cites Extending clip for category-to-image re- trieval in e-commerce.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Extending clip for category-to-image re- trieval in e-commerce

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:44.007546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:44.007546Z digest=sha256:b8f9374484bfcb0747176c803c9ef2b567133cfefb7ae485eb28cc12020e7f99

Observation 1d7566e4-c94a-455d-bd00-0ed6b9d6eacb · outbound

This paper cites Aligning AI with shared human values.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Aligning AI with shared human values

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:44.011271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:44.011271Z digest=sha256:08e5ce059963c2acecf67a80c5b670e760204fe97ba8e1c6f3fcdcf8da63e863

Observation 2c76dce9-fa76-4d1c-be3a-ab352787a12d · outbound

This paper cites Measuring massive multitask language BIBLIOGRAPHY 129 understanding.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Measuring massive multitask language BIBLIOGRAPHY 129 understanding

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:44.015165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:44.015165Z digest=sha256:46414fa798b06a088d78f9d641465a26e7e78c5f9f78dd54d02733cdc72b6f8e

Observation 1212c0c6-a34c-4ad7-af8c-050bec2fc8b1 · outbound

This paper cites Training Compute-Optimal Large Language Models.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Training Compute-Optimal Large Language Models

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:44.019604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:44.019604Z digest=sha256:497aa27cba210d48b63edfc4269ae75dc26d0675f8dd4be45e33479bc54f9ffd

Observation 89b04e8d-6a7f-4972-bd88-348b7ab06d90 · outbound

This paper cites ORPO: Monolithic Preference Optimization without Reference Model.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective ORPO: Monolithic Preference Optimization without Reference Model

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:44.024618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:44.024618Z digest=sha256:eba2c2c159d6e349dd17299aea006553448d41f0d7908e80fa1ce4bc276a9fde

Observation 387cd108-173b-488f-896a-efb847a5ec91 · outbound

This paper cites Unnatural in- structions: Tuning language models with (almost) no human labor.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Unnatural in- structions: Tuning language models with (almost) no human labor

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:44.028585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:44.028585Z digest=sha256:c3d20a91e567f8b01682b4b102c0c5cf4ab1a00f524aee04141107f3f83c454f

Observation 6591317c-b6fe-468d-bda6-ec2fe5272144 · outbound

This paper cites Human feedback is not gold standard.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Human feedback is not gold standard

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:44.032574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:44.032574Z digest=sha256:91b706e5df19bc715f1342d95b38586d6b74c7351fce1d13446c8e926fc21fbf

Observation d5c18acf-5c37-48fc-80c6-da3486230df3 · outbound

This paper cites Meta-learning the differ- ence: Preparing large language models for efficient adaptation.Transactions of the Association for Computational Linguistics, 10:1249–1265, 2022.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Meta-learning the differ- ence: Preparing large language models for efficient adaptation.Transactions of the Association for Computational Linguistics, 10:1249–1265, 2022

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:44.036748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:44.036748Z digest=sha256:c587d66425d261c654276a7df04a786fcf0477dd25f73dd4ae94ec5645fce0ef

Observation 972d2f27-166c-4da7-8008-730a3bac1ee1 · outbound

This paper cites Parameter-efficient transfer learning for NLP.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Parameter-efficient transfer learning for NLP

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:44.040532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:44.040532Z digest=sha256:6f01987bd8a510f779cdc706f8015d5d5d435411d582f697e05210d3088c18a1

Observation 87491cae-c3a0-4940-a3b9-7ac4efb15e71 · outbound

This paper cites Universal language model fine-tuning for text classification.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Universal language model fine-tuning for text classification

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:44.044640Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:44.044640Z digest=sha256:f5dac454e8ee1d70a004f34edd971d1b4ee6d13d7e99f5ecb7a0a27a25d756c6

Observation c159ef43-dec7-4a8d-a25c-16cd0d11e52c · outbound

This paper cites Can llms learn from a single exam- ple?, 2023.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Can llms learn from a single exam- ple?, 2023

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:44.049025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:44.049025Z digest=sha256:8fce6c63b00813d53808297f3eee45aebf1d1e481143ba698318092c89c9ddc2

Observation 070e5a04-10ca-489b-9f45-84d6e79b27f5 · outbound

This paper cites Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:44.053093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:44.053093Z digest=sha256:02d767930c6af5789b80d117897f4e7fb1fcb4c80383709727025548dcb85cc0

Observation 0fe4b3bd-abe6-48bc-8d52-5a324ea9b841 · outbound

This paper cites Mining and summarizing customer reviews.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Mining and summarizing customer reviews

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:44.057136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:44.057136Z digest=sha256:a12fde4dde62a408527ba33b500aa4393ed939cf747d7984eb1414a458deb00f

Observation f5bda0a5-dc72-4678-a121-4455a329f013 · outbound

This paper cites Instruction fine-tuning: Does prompt loss mat- ter?, 2024.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Instruction fine-tuning: Does prompt loss mat- ter?, 2024

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:44.061742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:44.061742Z digest=sha256:5fa94b293e1b4744d902bb2834e30ead4986a57d315c743113c47bca4fcccab8

Observation 8bb9cdef-5d9f-4a89-bffc-0f18e730a2b8 · outbound

This paper cites Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:44.066010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:44.066010Z digest=sha256:ba12324edb2f59db4283e6262dabbe840e1e358bc7a863b60fdf9279475ea4ca

Observation 8fc32d23-017b-4d41-aa13-4f3505d20fe2 · outbound

This paper cites Hyperdecoders: Instance-specific de- coders for multi-task NLP.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Hyperdecoders: Instance-specific de- coders for multi-task NLP

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:44.070392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:44.070392Z digest=sha256:ac78f479ecccf90cee9a0fbf932fb33897e10cf2a43914bce02e82a00eaea7de

Observation 03ee083b-576c-4b2a-9424-5ee6b357137d · outbound

This paper cites Camels in a changing climate: Enhancing lm adaptation with tulu 2, 2023.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Camels in a changing climate: Enhancing lm adaptation with tulu 2, 2023

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:44.074811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:44.074811Z digest=sha256:94ef4137748cd2cce761ed1fe600a16462ef598cad7297d9dc98a35070f2bea6

Observation d50d8371-1ccb-4c20-b3e7-6f0c7d166ac6 · outbound

This paper cites Bartoldson, Bhavya Kailkhura, Avi Schwarzschild, Aniruddha Saha, Micah Goldblum, Jonas Geiping, and Tom Goldstein.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Bartoldson, Bhavya Kailkhura, Avi Schwarzschild, Aniruddha Saha, Micah Goldblum, Jonas Geiping, and Tom Goldstein

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:44.079269Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:44.079269Z digest=sha256:b0e9d9d58e8d303930597bb4590874259c95109735430e091d6eabb260b4667e

Observation fa6bd133-d5a4-4ccc-ac10-fbff2c2ca7c9 · outbound

This paper cites Representation learning for grounded spatial reasoning.Transactions of the Association for Computational Linguistics, 6:49–61, 2018.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Representation learning for grounded spatial reasoning.Transactions of the Association for Computational Linguistics, 6:49–61, 2018

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:44.083308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:44.083308Z digest=sha256:7992f5433249a163402e37586a68e8a3ec6f81673089b212af11260a953b0b0d

Observation 452c341f-57e9-48ef-aed8-cb4f10f2dc73 · outbound

This paper cites LIMIT: Less Is More for Instruction Tuning Across Evaluation Paradigms.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective LIMIT: Less Is More for Instruction Tuning Across Evaluation Paradigms

Reference 99

Resolution
verified exact
local_arxiv, observed 2026-08-06T22:44:45.465717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:44:44.087462Z digest=sha256:ab5cc7dff15714d400a000c8bd52e8152e690dc2b421ea2e748b54646651d26d

Observation 013171b2-3b55-4642-b22c-d29254e49371 · outbound

This paper cites Scaling Laws for Neural Language Models.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Scaling Laws for Neural Language Models

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:44.091445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:44.091445Z digest=sha256:0962b7321212a809283429bf5cf2d2bd7c68cd6bd46049ba3d071f1c5c2e438f

Observation 58e56958-bd20-4f8c-aa66-aac08c5718fb · outbound

This paper cites Parameter-efficient multi-task fine-tuning for transformers via shared hypernetworks.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Parameter-efficient multi-task fine-tuning for transformers via shared hypernetworks

Reference 101

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:44.095688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:44.095688Z digest=sha256:4ed29e215521c1664d2adc0a6885b80f57dadbac6ef68ff2b0d8fec25cc17fd0

Observation df9ceb9b-f737-4228-b364-74eb76ee197b · outbound

This paper cites Looking beyond the surface: A challenge set for reading compre- hension over multiple sentences.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Looking beyond the surface: A challenge set for reading compre- hension over multiple sentences

Reference 102

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:44.099652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:44.099652Z digest=sha256:32bfadbfe6d059bd257a4c4c84b3e72d7b34a047d61115333fd047f79a51b683

Observation 418a7124-be8a-43ae-a98a-c33d2d53fdeb · outbound

This paper cites UNIFIEDQA: Crossing for- mat boundaries with a single QA system.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective UNIFIEDQA: Crossing for- mat boundaries with a single QA system

Reference 103

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:44.103686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:44.103686Z digest=sha256:f17e16fcf674bd23fcfc64a8e8b026e42d827c1bf72077971d4bdd8d0cfcfa08

Observation e69bab42-9f29-4524-b3f9-ec54819353d3 · outbound

This paper cites Scitail: A textual entail- ment dataset from science question answering.Proceedings of the AAAI Conference on Artificial Intelligence, 32(1), Apr.

Optimising Language Models for Downstream Tasks: A Post-Training Perspective Scitail: A textual entail- ment dataset from science question answering.Proceedings of the AAAI Conference on Artificial Intelligence, 32(1), Apr

Reference 104

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:44.107619Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:44.107619Z digest=sha256:a4bf85f1c327cf0ef513dc43135c381b036dc1703a5a4370050a83937b142128

Pith citing papers

No inbound Pith citation observations are available.