REVIEW 2 major objections 5 minor 177 references
Small Data Explainer -- The impact of small data methods in everyday life
T0 review · 2 major / 5 minor · reviewed 2026-08-06 · deepseek-v4-flash
Pith's one-line read The paper's central claim is that small data problems across applications share three recurring themes — similarity, transfer, and uncertainty — and that naming them gives statistics, computer science, and policy a shared language.
desk verdict A useful interdisciplinary explainer of small data, but its central 'similarity, transfer, uncertainty' framework rests more on the authors' own vignettes than on the claimed evidence synthesis. read the letter →
The pith
A machine-rendered reading of the paper's core claim, the machinery that carries it, and where it could break.
The reading
What carries the argument
The organizing object is the thematic triad of similarity, transfer, and uncertainty. Similarity means quantitatively comparing individuals or datasets to decide whether and how to combine information. Transfer means importing information from external sources — other datasets, pre-trained models such as foundation models, or expert knowledge encoded as equations or rules — to enrich a small dataset. Uncertainty means quantifying the reliability of estimates, predictions, and of the similarity and transfer steps themselves. The paper uses this triad both as a descriptive lens for case studies and as a proposed shared vocabulary that links knowledge-driven modelling from statistics and mathematics with data-driven modelling from computer science. It also contrasts two pictures of small data: as extreme values at the tails of one distribution, versus as a distinct subgroup with its own distribution, arguing the latter better captures underrepresented groups.
What would settle it
A systematic survey of real small-data studies, coded for whether their central analytic difficulties reduce to similarity, transfer, and uncertainty, would refute the framework if a substantial share of applications required different recurring sub-tasks.
Extended reading notes
Core claim
The paper's central claim is that small data is not a residual category defined only by what it lacks. Across four constructed scenarios — rare disease dosing, secure code generation with large language models, privacy-preserving financial recommendations, and wearable fall detection — the same three sub-tasks recur: assessing similarity, transferring information, and quantifying uncertainty. The authors argue that recognizing these recurring themes provides the missing shared language across disciplines, so that statisticians, computer scientists, and policymakers can recognise common problems, reuse each other's methods, and build a joint research agenda. If this is right, small data methods that already work in fields such as rare disease or precision medicine can be transplanted to less mature fields, and the assumption that bigger data is always better can be deliberately set aside.
Load-bearing premise
The load-bearing premise is that the four illustrative scenarios the authors constructed are representative of small data challenges; if real applications surface additional or different recurring themes, the framework would be incomplete.
Editorial extensions
If this is right
- If similarity, transfer, and uncertainty are the common sub-tasks, then a method developed for one small data field, such as rare disease, can be ported to another, such as wearable health, once the shared sub-task is identified.
- Small data approaches can make AI and policy decisions more inclusive by modelling underrepresented groups as their own distributions rather than as outliers of a majority distribution.
- Foundation models can serve as a transfer channel: pre-trained knowledge is adapted to a small target dataset via fine-tuning or in-context learning, with the caveat that similarity must be assessed to avoid harmful transfer such as hallucination.
- Data minimisation and small data analysis are compatible when uncertainty is quantified, allowing privacy-preserving, on-device personalisation.
- Policies that encourage open data-sharing standards and external validation would help small data models be checked against similar datasets.
Reading between the lines
- The paper does not test its triad against a systematic sample; a natural next step would be coding published small-data applications by theme and checking whether methods cluster by theme across domains.
- The triad can be read as a design checklist: before building a small-data system, specify how similarity is measured, what information is transferred, and how uncertainty is reported; the paper's framework implies this would improve interdisciplinary collaboration.
- The framework also suggests a policy direction the authors only gesture at: funders could prioritise small, high-quality datasets plus external knowledge over ever-larger collection, which would align with privacy and cost goals.
Editorial analysis
A structured set of objections, weighed in public.
Referee Report
Summary. This paper offers a conceptual, cross-disciplinary overview of small data: it contrasts small data with big data, discusses the limitations of big data for under-represented groups, introduces four hypothetical vignettes (Boxes 1-4) to motivate three recurring themes (similarity, transfer, uncertainty), surveys several application fields (rare disease, precision medicine, assistive technology, data minimisation, generative AI), and then provides a technical tour of methods such as foundation models, few-shot learning, meta-learning, and knowledge-driven modelling. The stated goal is to foster a shared language for small data research across statistics, computer science, and policy. The central claim is that the three themes can structure the field of small data.
Significance. If the three-theme framework is accepted, it could indeed serve as a useful bridge between disciplines: the paper links concrete applications to specific methodological families and makes a plausible case that similarity, transfer, and uncertainty are productive lenses for organizing small data methods. The paper is clearly written, contains a useful glossary, and draws on a broad reference base that includes both statistical and computer science literatures. However, the central empirical claim—that these three themes are the recurring common challenges of small data—is not independently evidenced: it is derived from the paper's own hypothetical examples, which are themselves constructed around those themes. The paper's value is therefore stronger as a proposed conceptual framework and pedagogical explainer than as an evidence-based synthesis, and the framing should be adjusted accordingly.
major comments (2)
- [Section 2.4] The central claim that the three themes 'repeatedly emerged' is supported only by the sentence 'In the examples given above (in particular Boxes 1-4), there were three themes that repeatedly emerged: similarity, transfer, and uncertainty.' Boxes 1-4 are, however, self-authored hypothetical scenarios explicitly labeled 'Hypothetical exemplary scenario,' and each was visibly constructed around the themes: Box 1 discusses matching similar patients and transferring information across age groups, Box 2 discusses a similarity metric, fine-tuning/transfer, and confidence warnings, Box 3 discusses similarity measures, transferring generic models, and uncertainty evaluations, and Box 4 discusses subgroups/similarity, information transfer, and uncertainty. Finding the three themes in examples designed to illustrate them is circular and does not establish that they are the recurring themes of the field. To make this claim load-bearing, the paper must either add an independent evidence base (e.g., a systematic corpus of real small data applications with explicit selection criteria and a documented synthesis method) or reframe the three themes as a proposed organizing framework rather than an empirically observed regularity. The conclusion in Section 5, which elevates the three themes to a basis for structuring the field and a shared language, inherits this problem.
- [Section 1] The Introduction states that the paper's three primary questions were answered 'through extensive desk review and evidence synthesis,' but no review protocol, corpus, inclusion criteria, or synthesis method is described anywhere in the manuscript. An 'extensive desk review' is claimed, yet the reader cannot verify what was reviewed or how themes were synthesized. In a revised version the authors should either add a brief methods appendix describing the review and synthesis process or replace this claim with a more modest statement, such as 'we illustrate the themes through selected examples and application fields.' Without this adjustment, the paper's empirical grounding is not verifiable and the unsupported 'repeatedly emerged' claim in Section 2.4 is further exposed.
minor comments (5)
- [Section 1] The last paragraph of Section 1 reads 'propose an agenda for increasing the use of small data in in Section 5'; the duplicated 'in' should be removed.
- [Section 3.5] The phrase 'as illustrated in 2' is missing a figure reference; it should read 'as illustrated in Figure 1' or whichever figure is intended.
- [Glossary after Section 2.2] The glossary entry for foundation models contains the typo 'These capabilites' and should read 'capabilities.'
- [Several sections] The word 'Yet' is rendered as 'Y et' in multiple places (e.g., Section 2.4.2, Section 3.4, Section 4.8); these spacing errors should be corrected.
- [Reference [92]] Reference [92] lists the surnames of historical scientists—'J. C. M. Darwin, G. Mendel, R. Franklin, J. Watson, et al.'—as authors of a Nature Biomedical Engineering paper. This appears to be an erroneous or fabricated citation and must be verified and corrected.
Circularity Check
The three-theme framework is derived from self-authored vignettes that were constructed around those themes.
-
self definitional
[Section 2.4 (Common themes: similarity, transfer, and uncertainty)]
"In the examples given above (in particular Boxes 1-4), there were three themes that repeatedly emerged: similarity, transfer, and uncertainty."
Boxes 1-4 are 'Hypothetical exemplary scenario[s]' authored by the paper, and each is worded to include the target themes: Box 1 mentions matching similar patients, transfer across age groups, and difficult matching; Box 2 mentions a similarity metric, fine-tuning, and confidence warnings; Box 3 mentions similarity measures, transferring a generic model, and uncertainty evaluations; Box 4 mentions similarity, information transfer, and uncertainty. Section 1 also says the same section 'introduces the themes of similarity, transfer, and uncertainty.' So the 'repeated emergence' is a design feature, not independent evidence; Section 5 then calls these the 'common themes' for structuring the field, making the derivation circular.
full rationale
The paper's central, load-bearing claim is that similarity, transfer, and uncertainty are the 'common themes' that structure small data (Sections 2.4 and 5). The evidence offered is that these themes 'repeatedly emerged' from Boxes 1-4. Those boxes are hypothetical vignettes written by the authors, and each is explicitly worded to include all three themes; Section 1 also states that the section 'introduces the themes of similarity, transfer, and uncertainty' right where the boxes appear. Finding the target themes in examples constructed around them is self-definitional: the recurrence is guaranteed by the design. The paper's stated method, 'extensive desk review and evidence synthesis', is not documented with any protocol, corpus, or selection criteria, so it cannot independently corroborate the recurrence. The self-citations in the paper (e.g., [96], [102], [103], [113], [159]) are used only as examples of methods and are not load-bearing for the framework, so they do not add circularity. The themes themselves do have independent content in established fields—statistical matching, transfer learning, and uncertainty quantification—so the framework is not wholly empty. The circularity is thus partial: the central claim's derivation from the examples reduces to the authors' own construction, but the framework also stands on independent method families.
Assumptions & free parameters
assumptions (2)
- domain assumption Small data settings are those with limited information, and the three themes similarity, transfer, and uncertainty are the common challenges across these settings.
- ad hoc to paper The hypothetical case studies in Boxes 1-4 are representative of small data applications.
Cite this review
Pith. "Pith review of Small Data Explainer -- The impact of small data methods in everyday life." pith.science (2026). https://pith.science/paper/GBHY7FVN
@misc{pith2026250711773,
author = {Pith},
title = {Pith review of: Small Data Explainer -- The impact of small data methods in everyday life},
year = {2026},
howpublished = {\url{https://pith.science/paper/GBHY7FVN}},
note = {Machine review of arXiv:2507.11773}
}
read the original abstract
The emergence of breakthrough artificial intelligence (AI) techniques has led to a renewed focus on how small data settings, i.e., settings with limited information, can benefit from such developments. This includes societal issues such as how best to include under-represented groups in data-driven policy and decision making, or the health benefits of assistive technologies such as wearables. We provide a conceptual overview, in particular contrasting small data with big data, and identify common themes from exemplary case studies and application areas. Potential solutions are described in a more detailed technical overview of current data analysis and modelling techniques, highlighting contributions from different disciplines, such as knowledge-driven modelling from statistics and data-driven modelling from computer science. By linking application settings, conceptual contributions and specific techniques, we highlight what is already feasible and suggest what an agenda for fully leveraging small data might look like.
Figures
Reference graph
Works this paper leans on
-
[1]
When small data beats big data,
J. J. Faraway and N. H. Augustin, “When small data beats big data,” Statistics & Probability Letters, vol. 136, pp. 142–145, 2018
2018
-
[2]
Small data’s big AI potential,
H. Chahal, H. Toner, and I. Rahkovsky, “Small data’s big AI potential,” Center for Security and Emerging Technology, 2021
2021
-
[3]
Beyond big data,
S. Ehrenberg-Silies, “Beyond big data,” Büro für Technikfolgen-Abschätzung beim Deutschen Bundestag (TAB), 2019
2019
-
[4]
Andrew Ng, AI minimalist: The machine-learning pioneer says small is the new big,
E. Strickland, “Andrew Ng, AI minimalist: The machine-learning pioneer says small is the new big,” IEEE Spectrum, vol. 59, no. 4, pp. 22–50, 2022
2022
-
[5]
Breaking the data barrier: A review of deep learning techniques for democratizing AI with small datasets,
I. H. Rather, S. Kumar, and A. H. Gandomi, “Breaking the data barrier: A review of deep learning techniques for democratizing AI with small datasets,” Artificial Intelligence Review, vol. 57, no. 9, p. 226, 2024
2024
-
[6]
Big data from small data: Data-sharing in the ‘long tail’ of neuroscience,
A. R. Ferguson, J. L. Nielson, M. H. Cragin, A. E. Bandrowski, and M. E. Martone, “Big data from small data: Data-sharing in the ‘long tail’ of neuroscience,” Nature Neuroscience, vol. 17, no. 11, pp. 1442–1447, 2014
2014
-
[7]
Why we need a small data paradigm,
E. B. Hekler, P . Klasnja, G. Chevance, N. M. Golaszewski, D. Lewis, and I. Sim, “Why we need a small data paradigm,” BMC Medicine, vol. 17, pp. 1–9, 2019
2019
-
[8]
Axes of a revolution: Challenges and promises of big data in healthcare,
S. Shilo, H. Rossman, and E. Segal, “Axes of a revolution: Challenges and promises of big data in healthcare,” Nature Medicine, vol. 26, no. 1, pp. 29–38, 2020
2020
Show all 177 references
-
[9]
Data—from objects to assets,
S. Leonelli, “Data—from objects to assets,” Nature, vol. 574, no. 7778, pp. 317–320, 2019
2019
-
[10]
Learning with small data,
Z. Li, H. Y ao, and F . Ma, “Learning with small data,” in International Conference on Web Search and Data Mining, 2020, pp. 884–887
2020
-
[11]
Handling limited datasets with neural networks in medical applications: A small-data approach,
T. Shaikhina and N. A. Khovanova, “Handling limited datasets with neural networks in medical applications: A small-data approach,” Artificial Intelligence in Medicine , vol. 75, pp. 51–63, 2017
2017
-
[12]
Machine learning methods for small data challenges in molecular science,
B. Dou, Z. Zhu, E. Merkurjev, L. Ke, L. Chen, J. Jiang, et al., “Machine learning methods for small data challenges in molecular science,” Chemical Reviews, vol. 123, no. 13, pp. 8736– 8780, 2023
2023
-
[13]
Machine learning on small size samples: A synthetic knowledge synthesis,
P . Kokol, M. Kokol, and S. Zagoranski, “Machine learning on small size samples: A synthetic knowledge synthesis,” Science Progress, vol. 105, no. 1, 2022
2022
-
[14]
Artificial intelligence and statistics: Just the old wine in new wineskins?
L. Faes, D. A. Sim, M. van Smeden, U. Held, P . M. Bossuyt, and L. M. Bachmann, “Artificial intelligence and statistics: Just the old wine in new wineskins?” Frontiers in Digital Health , vol. 4, p. 833 912, 2022
2022
-
[15]
Small data, big decisions: Model selection in the small-data regime,
J. Bornschein, F . Visin, and S. Osindero, “Small data, big decisions: Model selection in the small-data regime,” 2020. arXiv: 2009.12583, pre-published
2020 arXiv
-
[16]
Trending: The promises and the challenges of big social data,
L. Manovich, “Trending: The promises and the challenges of big social data,” in Debates in the Digital Humanities, University of Minnesota Press, 2012, pp. 460–475. 26
2012
-
[17]
Critical questions for big data: Provocations for a cultural, techno- logical, and scholarly phenomenon,
D. Boyd and K. Crawford, “Critical questions for big data: Provocations for a cultural, techno- logical, and scholarly phenomenon,”Information, Communication and Society, vol. 15, no. 5, pp. 662–679, 2012
2012
-
[18]
Mayer-Schönberger and K
V. Mayer-Schönberger and K. Cukier, Big Data: A Revolution That Will Transform How We Live, Work and Think. London: Murray, 2013, 242 pp
2013
-
[19]
Big data in big companies,
T. H. Davenport and J. Dyché, “Big data in big companies,” International Institute for Analyt- ics, 2013. [Online]. Available: https://www.iqpc.com/media/7863/11710.pdf
2013
-
[20]
Here’s how technology has changed the world since 2000,
M. Hillyer. “Here’s how technology has changed the world since 2000,” World Economic Forum. (2020), [Online]. Available: https://www.weforum.org/stories/2020/11/heres- how-technology-has-changed-and-changed-us-over-the-past-20-years/
2020
-
[21]
Towards precision medicine,
E. A. Ashley, “Towards precision medicine,” Nature Reviews Genetics, vol. 17, no. 9, pp. 507– 522, 2016
2016
-
[22]
Recent advances and trends in predictive manufacturing systems in big data environment,
J. Lee, E. Lapira, B. Bagheri, and H.-a. Kao, “Recent advances and trends in predictive manufacturing systems in big data environment,”Manufacturing Letters, vol. 1, no. 1, pp. 38– 41, 2013
2013
-
[23]
A re- view of Gaucher disease pathophysiology, clinical presentation and treatments,
J. Stirnemann, N. Belmatoug, F . Camou, C. Serratrice, R. Froissart, C. Caillaud, et al., “A re- view of Gaucher disease pathophysiology, clinical presentation and treatments,”International Journal of Molecular Sciences, vol. 18, no. 2, p. 441, 2017
2017
-
[24]
Individualization of long-term enzyme replacement therapy for Gaucher disease,
H. C. Andersson, J. Charrow, P . Kaplan, P . Mistry, G. M. Pastores, A. Prakesh-Cheng,et al., “Individualization of long-term enzyme replacement therapy for Gaucher disease,” Genetics in Medicine, vol. 7, no. 2, pp. 105–110, 2005
2005
-
[25]
The value of being different,
J. Treviranus, “The value of being different,” in International Web for All Conference , 2019, pp. 1–7
2019
-
[26]
Adolphe Quetelet and the legacy of the “average man
D. Tafreshi, “Adolphe Quetelet and the legacy of the “average man” in psychology.,” History of Psychology, vol. 25, no. 1, p. 34, 2022
2022
-
[27]
People are people, but technology is not technol- ogy,
G. Marsden, A. Maunder, and M. Parker, “People are people, but technology is not technol- ogy,” Philosophical Transactions of the Royal Society A: Mathematical, Physical and Engi- neering Sciences, vol. 366, no. 1881, pp. 3795–3804, 2008
2008
-
[28]
G. J. Downey, Closed Captioning: Subtitling, Stenography, and the Digital Convergence of Text with Television. Baltimore: Johns Hopkins University Press, 2008, 387 pp
2008
-
[29]
Steinfeld and J
E. Steinfeld and J. Maisel, Universal Design: Creating Inclusive Environments. Hoboken, NJ: John Wiley & Sons, Inc, 2012, 382 pp
2012
-
[30]
Springer London, 2019
Web Accessibility: A Foundation for Research. Springer London, 2019
2019
-
[31]
“Accessibility came by accident
A. Pradhan, K. Mehta, and L. Findlater, ““Accessibility came by accident”: Use of voice- controlled intelligent personal assistants by people with disabilities,” inConference on Human Factors in Computing Systems, Association for Computing Machinery, 2018, pp. 1–13
2018
-
[32]
Disability, locative media, and complex ubiquity,
K. Ellis and G. Goggin, “Disability, locative media, and complex ubiquity,” in Ubiquitous Com- puting, Complexity and Culture, Routledge, 2015, pp. 272–287. 27
2015
-
[33]
Matching networks for one shot learning,
O. Vinyals, C. Blundell, T. Lillicrap, K. Kavukcuoglu, and D. Wierstra, “Matching networks for one shot learning,” in Advances in Neural Information Processing Systems , 2016. [Online]. Available: https://proceedings.neurips.cc/paper_files/paper/2016/hash/90e1357 833654983612f...
2016
-
[34]
On the opportunities and risks of foundation models,
R. Bommasani, D. A. Hudson, E. Adeli, R. Altman, S. Arora, S. von Arx, et al. , “On the opportunities and risks of foundation models,” 2022. arXiv: 2108.07258, pre-published
2022 arXiv
-
[35]
Back to the building blocks: A path toward secure and measurable software,
“Back to the building blocks: A path toward secure and measurable software,” White House Office of the National Cyber Director, 2024. [Online]. Available: https://bidenwhitehouse .archives.gov/wp-content/uploads/2024/02/Final-ONCD-Technical-Report.pdf
2024
-
[36]
The case for memory safety roadmaps,
Department of Homeland Security, “The case for memory safety roadmaps,” Cybersecurity Infrastructure and Security Agency, 2023. [Online]. Available: https://www.cisa.gov/sit es/default/files/2023-12/The-Case-for-Memory-Safe-Roadmaps-508c.pdf
2023
-
[37]
How secure is code generated by ChatGPT?,
R. Khoury, A. R. Avila, J. Brunelle, and B. M. Camara, “How secure is code generated by ChatGPT?,” 2023. arXiv: 2304.09655, pre-published
2023 arXiv
-
[38]
Data security and consumer trust in FinTech innovation in ger- many,
H. Stewart and J. Jürjens, “Data security and consumer trust in FinTech innovation in ger- many,”Information & Computer Security, vol. 26, no. 1, pp. 109–128, 2018
2018
-
[39]
Robo-what?, robo-why?, robo-how? – a systematic literature review of robo-advice,
A. Torno, D. R. Metzler, and V. Torno, “Robo-what?, robo-why?, robo-how? – a systematic literature review of robo-advice,”Pacific Asia Conference on Information Systems, 2021. [On- line]. Available: https://aisel.aisnet.org/pacis2021/92
2021
-
[40]
Disruptions and digital banking trends,
L. Wewege, J. Lee, and M. C. Thomsett, “Disruptions and digital banking trends,” Journal of Applied Finance & Banking, vol. 10, no. 6, pp. 1–2, 2020. [Online]. Available:https://idea s.repec.org//a/spt/apfiba/v10y2020i6f10_6_2.html
2020
-
[41]
The rise of consumer health wearables: Promises and barriers,
L. Piwek, D. A. Ellis, S. Andrews, and A. Joinson, “The rise of consumer health wearables: Promises and barriers,” PLoS medicine, vol. 13, no. 2, 2016
2016
-
[42]
Thousand Oaks, California: SAGE Publications, 2003
Handbook of Mixed Methods in Social and Behavioral Research. Thousand Oaks, California: SAGE Publications, 2003
2003
-
[43]
A survey on recent advances in wearable fall detection systems,
A. Ramachandran and A. Karuppiah, “A survey on recent advances in wearable fall detection systems,” BioMed Research International, vol. 2020, pp. 1–17, 2020
2020
-
[44]
Challenges, issues and trends in fall detection systems,
R. Igual, C. Medrano, and I. Plaza, “Challenges, issues and trends in fall detection systems,” BioMedical Engineering OnLine, vol. 12, no. 1, p. 66, 2013
2013
-
[45]
The effect of personalization on smartphone-based fall detectors,
C. Medrano, I. Plaza, R. Igual, A. Sánchez, and M. Castro, “The effect of personalization on smartphone-based fall detectors,” Sensors, vol. 16, no. 1, p. 117, 2016
2016
-
[46]
UP-Fall detection dataset: A multimodal approach,
L. Martínez-Villaseñor, H. Ponce, J. Brieva, E. Moya-Albor, J. Núñez-Martínez, and C. Peñafort- Asturiano, “UP-Fall detection dataset: A multimodal approach,”Sensors, vol. 19, no. 9, p. 1988, 2019
1988
-
[47]
Aleatoric and epistemic uncertainty in machine learning: An introduction to concepts and methods,
E. Hüllermeier and W. Waegeman, “Aleatoric and epistemic uncertainty in machine learning: An introduction to concepts and methods,” Machine Learning, vol. 110, no. 3, pp. 457–506, 2021. 28
2021
-
[48]
Washington (DC): National Academies Press (US), 2011
National Research Council (US), Toward Precision Medicine: Building a Knowledge Net- work for Biomedical Research and a New Taxonomy of Disease. Washington (DC): National Academies Press (US), 2011. [Online]. Available: http://www.ncbi.nlm.nih.gov/books /NBK91503/
2011
-
[49]
Ciardiello, D
F . Ciardiello, D. Arnold, P . G. Casali, A. Cervantes, J.-Y . Douillard, A. Eggermont,et al., “De- livering precision medicine in oncology today and in future — the promise and challenges of personalised cancer medicine: A position paper by the european society for medical on...
2014
-
[50]
Precision medicine in rare diseases,
I. Villalón-García, M. Álvarez-Córdoba, J. M. Suárez-Rivero, S. Povea-Cabello, M. Talaverón- Rey, A. Suárez-Carrillo, et al., “Precision medicine in rare diseases,” Diseases, vol. 8, no. 4, p. 42, 2020
2020
-
[51]
Machine learning in rare disease,
J. Banerjee, J. N. Taroni, R. J. Allaway, D. V. Prasad, J. Guinney, and C. Greene, “Machine learning in rare disease,” Nature Methods, vol. 20, no. 6, pp. 803–814, 2023
2023
-
[52]
The state of the art and future opportunities for using longitudinal n-of-1 methods in health behaviour re- search: A systematic literature overview,
S. McDonald, F . Quinn, R. Vieira, N. O’Brien, M. White, D. W. Johnston, et al., “The state of the art and future opportunities for using longitudinal n-of-1 methods in health behaviour re- search: A systematic literature overview,”Health Psychology Review, vol. 11, no. 4, pp....
2017
-
[53]
The history and development of n-of-1 trials,
R. Mirza, S. Punja, S. Vohra, and G. Guyatt, “The history and development of n-of-1 trials,” Journal of the Royal Society of Medicine, vol. 110, no. 8, pp. 330–340, 2017
2017
-
[54]
N-of-1 trials in the medical literature: A systematic review,
N. B. Gabler, N. Duan, S. Vohra, and R. L. Kravitz, “N-of-1 trials in the medical literature: A systematic review,”Medical Care, vol. 49, no. 8, pp. 761–768, 2011
2011
-
[55]
Individual (n-of-1) trials can be combined to give population comparative treatment effect estimates: Methodologic considerations,
D. R. Zucker, R. Ruthazer, and C. H. Schmid, “Individual (n-of-1) trials can be combined to give population comparative treatment effect estimates: Methodologic considerations,”Jour- nal of Clinical Epidemiology, vol. 63, no. 12, pp. 1312–1323, 2010
2010
-
[56]
Precision oncology medicine: The clinical relevance of patient-specific biomarkers used to optimize cancer treatment,
K. T. Schmidt, C. H. Chau, D. K. Price, and W. D. Figg, “Precision oncology medicine: The clinical relevance of patient-specific biomarkers used to optimize cancer treatment,” The Journal of Clinical Pharmacology, vol. 56, no. 12, pp. 1484–1499, 2016
2016
-
[57]
Evidence for personalised medicine: Mecha- nisms, correlation, and new kinds of black box,
M. J. Walker, J. Bourke, and K. Hutchison, “Evidence for personalised medicine: Mecha- nisms, correlation, and new kinds of black box,” Theoretical Medicine and Bioethics, vol. 40, pp. 103–121, 2019
2019
-
[58]
Case-based reasoning: Foundational issues, methodological vari- ations, and system approaches,
A. Aamodt and E. Plaza, “Case-based reasoning: Foundational issues, methodological vari- ations, and system approaches,” AI Communications, vol. 7, no. 1, pp. 39–59, 1994
1994
-
[59]
Optimizing clinical practice with case-based reasoning approach,
C. Dussart, P . Pommier, V. Siranyan, G. Grelaud, and S. Dussart, “Optimizing clinical practice with case-based reasoning approach,” Journal of Evaluation in Clinical Practice , vol. 14, no. 5, pp. 718–720, 2008
2008
-
[60]
Similarity measures for retrieval in case-based rea- soning systems,
T. W. Liao, Z. Zhang, and C. R. Mount, “Similarity measures for retrieval in case-based rea- soning systems,” Applied Artificial Intelligence, vol. 12, no. 4, pp. 267–288, 1998. 29
1998
-
[61]
Fuzzy set theory and uncertainty in case-based reasoning,
R. Weber, “Fuzzy set theory and uncertainty in case-based reasoning,” Engineering Intelli- gent Systems for Electrical Engineering and Communications, vol. 14, no. 3, p. 121, 2007
2007
-
[62]
Multiple retrieval case-based reasoning for incomplete datasets,
N. Löw, J. Hesser, and M. Blessing, “Multiple retrieval case-based reasoning for incomplete datasets,” Journal of Biomedical Informatics, vol. 92, p. 103 127, 2019
2019
-
[63]
Case-based reasoning in transfer learning,
D. W. Aha, M. Molineaux, and G. Sukthankar, “Case-based reasoning in transfer learning,” in Case-Based Reasoning Research and Development, Springer, 2009, pp. 29–44
2009
-
[64]
Wearable and miniaturized sensor technologies for personal- ized and preventive medicine,
A. Tricoli, N. Nasiri, and S. De, “Wearable and miniaturized sensor technologies for personal- ized and preventive medicine,” Advanced Functional Materials, vol. 27, no. 15, p. 1 605 271, 2017
2017
-
[65]
Wearable health devices in health care: Narrative systematic review,
L. Lu, J. Zhang, Y . Xie, F . Gao, S. Xu, X. Wu, et al., “Wearable health devices in health care: Narrative systematic review,”JMIR mHealth and uHealth, vol. 8, no. 11, e18907, 2020
2020
-
[66]
How self-tracking and the quantified self promote health and well-being: Systematic review,
S. Feng, M. ki, A. Dhir, and H. Salmela, “How self-tracking and the quantified self promote health and well-being: Systematic review,” Journal of Medical Internet Research , vol. 23, no. 9, e25171, 2021
2021
-
[67]
Adaptive extreme edge computing for wearable devices,
E. Covi, E. Donati, X. Liang, D. Kappel, H. Heidari, M. Payvand, et al., “Adaptive extreme edge computing for wearable devices,”Frontiers in Neuroscience, vol. 15, 2021
2021
-
[68]
Convergence of edge computing and deep learning: A comprehensive survey,
X. Wang, Y . Han, V. C. M. Leung, D. Niyato, X. Y an, and X. Chen, “Convergence of edge computing and deep learning: A comprehensive survey,” IEEE Communications Surveys & Tutorials, vol. 22, no. 2, pp. 869–904, 2020
2020
-
[69]
Deep learning and model personaliza- tion in sensor-based human activity recognition,
A. Ferrari, D. Micucci, M. Mobilio, and P . Napoletano, “Deep learning and model personaliza- tion in sensor-based human activity recognition,”Journal of Reliable Intelligent Environments, vol. 9, no. 1, pp. 27–39, 2023
2023
-
[70]
[Online]
European Parliament, Regulation (EU) 2016/679 of the European Parliament and of the Council of 27 April 2016 on the protection of natural persons with regard to the processing of personal data and on the free movement of such data, and repealing Directive 95/46/EC (General Dat...
2016
-
[71]
Ex- plaining automated decision-making: A multinational study of the GDPR right to meaningful information,
J. Dexe, U. Franke, K. Söderlund, N. van Berkel, R. H. Jensen, N. Lepinkäinen, et al., “Ex- plaining automated decision-making: A multinational study of the GDPR right to meaningful information,” The Geneva Papers on Risk and Insurance - Issues and Practice, vol. 47, no. 3, pp...
2022
-
[72]
Can I opt out yet?: GDPR and the global illusion of cookie control,
I. Sanchez-Rola, M. Dell’Amico, P . Kotzias, D. Balzarotti, L. Bilge, P .-A. Vervier, et al., “Can I opt out yet?: GDPR and the global illusion of cookie control,” in ACM Asia Conference on Computer and Communications Security, 2019, pp. 340–351
2019
-
[73]
Private traits and attributes are predictable from digital records of human behavior,
M. Kosinski, D. Stillwell, and T. Graepel, “Private traits and attributes are predictable from digital records of human behavior,”National Academy of Sciences, vol. 110, no. 15, pp. 5802– 5805, 2013. 30
2013
-
[74]
Zero-shot text-to- image generation,
A. Ramesh, M. Pavlov, G. Goh, S. Gray, C. Voss, A. Radford, et al. , “Zero-shot text-to- image generation,” in International Conference on Machine Learning, 2021, pp. 8821–8831. [Online]. Available: https://proceedings.mlr.press/v139/ramesh21a.html
2021
-
[75]
Hierarchical text-conditional image generation with CLIP latents,
A. Ramesh, P . Dhariwal, A. Nichol, C. Chu, and M. Chen, “Hierarchical text-conditional image generation with CLIP latents,” 2022. arXiv: 2204.06125, pre-published
2022 arXiv
-
[76]
Pop music transformer: Beat-based modeling and generation of expressive pop piano compositions,
Y .-S. Huang and Y .-H. Y ang, “Pop music transformer: Beat-based modeling and generation of expressive pop piano compositions,” in ACM International Conference on Multimedia, 2020, pp. 1180–1188
2020
-
[77]
UniAudio: An audio foundation model toward universal audio generation,
D. Y ang, J. Tian, X. Tan, R. Huang, S. Liu, X. Chang, et al., “UniAudio: An audio foundation model toward universal audio generation,” 2024. arXiv: 2310.00704, pre-published
2024 arXiv
-
[78]
Visual ChatGPT: Talking, drawing and editing with visual foundation models,
C. Wu, S. Yin, W. Qi, X. Wang, Z. Tang, and N. Duan, “Visual ChatGPT: Talking, drawing and editing with visual foundation models,” 2023. arXiv: 2303.04671, pre-published
2023 arXiv
-
[79]
mPLUG-2: A modularized multi-modal foundation model across text, image and video,
H. Xu, Q. Y e, M. Y an, Y . Shi, J. Y e, Y . Xu, et al. , “mPLUG-2: A modularized multi-modal foundation model across text, image and video,” in International Conference on Machine Learning, 2023, pp. 38 728–38 748. [Online]. Available:https://proceedings.mlr.press /v202/xu23s.html
2023
-
[80]
Large-scale text-to-image genera- tion models for visual artists’ creative works,
H.-K. Ko, G. Park, H. Jeon, J. Jo, J. Kim, and J. Seo, “Large-scale text-to-image genera- tion models for visual artists’ creative works,” in International Conference on Intelligent User Interfaces, 2023, pp. 919–933
2023
-
[81]
Assessing GPT -4 for cell type annotation in single-cell RNA-seq analysis,
W. Hou and Z. Ji, “Assessing GPT -4 for cell type annotation in single-cell RNA-seq analysis,” Nature Methods, pp. 1–4, 2024
2024
-
[82]
Large-language models facilitate discovery of the molecular signatures regulating sleep and activity,
D. Peng, L. Zheng, D. Liu, C. Han, X. Wang, Y . Y ang,et al., “Large-language models facilitate discovery of the molecular signatures regulating sleep and activity,”Nature Communications, vol. 15, no. 1, p. 3685, 2024
2024
-
[83]
Bidirectional generation of structure and properties through a single molecular foundation model,
J. Chang and J. C. Y e, “Bidirectional generation of structure and properties through a single molecular foundation model,” Nature Communications, vol. 15, no. 1, p. 2323, 2024
2024
-
[84]
Sources of hallucination by large language models on inference tasks,
N. McKenna, T. Li, L. Cheng, M. J. Hosseini, M. Johnson, and M. Steedman, “Sources of hallucination by large language models on inference tasks,” 2023. arXiv: 2305.14552, pre- published
2023 arXiv
-
[85]
Language models are few-shot learners,
T. B. Brown, B. Mann, N. Ryder, M. Subbiah, J. Kaplan, P . Dhariwal,et al., “Language models are few-shot learners,” 2020. arXiv: 2005.14165, pre-published
2020 arXiv
-
[86]
Opportunities and challenges in modeling emerging infec- tious diseases,
C. J. E. Metcalf and J. Lessler, “Opportunities and challenges in modeling emerging infec- tious diseases,” Science, vol. 357, no. 6347, pp. 149–152, 2017
2017
-
[87]
A cellular automaton model for tumour growth in inhomogeneous environment,
T. Alarcón, H. M. Byrne, and P . K. Maini, “A cellular automaton model for tumour growth in inhomogeneous environment,” Journal of Theoretical Biology , vol. 225, no. 2, pp. 257–274, 2003. 31
2003
-
[88]
Is there a role for statistics in artificial intelligence?
S. Friedrich, G. Antes, S. Behr, H. Binder, W. Brannath, F . Dumpert, et al., “Is there a role for statistics in artificial intelligence?” Advances in Data Analysis and Classification, vol. 16, pp. 823–846, 2022
2022
-
[89]
Medical domain knowledge in domain-agnostic generative AI,
J. N. Kather, N. Ghaffari Laleh, S. Foersch, and D. Truhn, “Medical domain knowledge in domain-agnostic generative AI,” npj Digital Medicine, vol. 5, no. 1, pp. 1–5, 2022
2022
-
[90]
A review of some techniques for in- clusion of domain-knowledge into deep neural networks,
T. Dash, S. Chitlangia, A. Ahuja, and A. Srinivasan, “A review of some techniques for in- clusion of domain-knowledge into deep neural networks,” Scientific Reports, vol. 12, no. 1, p. 1040, 2022
2022
-
[91]
Data and domain knowledge dual-driven artificial intelligence: Survey, applications, and challenges,
J. Nie, J. Jiang, Y . Li, H. Wang, S. Ercisli, and L. Lv, “Data and domain knowledge dual-driven artificial intelligence: Survey, applications, and challenges,” Expert Systems, vol. 42, no. 1, e13425, 2025
2025
-
[92]
Pooling the strengths of data and models,
J. C. M. Darwin, G. Mendel, R. Franklin, J. Watson, et al., “Pooling the strengths of data and models,” Nature Biomedical Engineering, vol. 5, no. 4, pp. 291–292, 2021
2021
-
[93]
The performance of different propensity score methods for estimating marginal odds ratios,
P . C. Austin, “The performance of different propensity score methods for estimating marginal odds ratios,” Statistics in Medicine, vol. 26, no. 16, pp. 3078–3094, 2007
2007
-
[94]
Moving towards best practice when using inverse probability of treatment weighting (IPTW) using the propensity score to estimate causal treatment effects in observational studies,
P . C. Austin and E. A. Stuart, “Moving towards best practice when using inverse probability of treatment weighting (IPTW) using the propensity score to estimate causal treatment effects in observational studies,” Statistics in Medicine, vol. 34, no. 28, pp. 3661–3679, 2015
2015
-
[95]
Methods for the analysis of sampled cohort data in the cox proportional hazards model,
O. Borgan, L. Goldstein, and B. Langholz, “Methods for the analysis of sampled cohort data in the cox proportional hazards model,” Annals of Statistics, vol. 23, no. 5, pp. 1749–1778, 1995
1995
-
[96]
A weighting approach for judging the effect of patient strata on high-dimensional risk prediction signatures,
V. Weyer and H. Binder, “A weighting approach for judging the effect of patient strata on high-dimensional risk prediction signatures,” BMC Bioinformatics, vol. 16, p. 294, 2015
2015
-
[97]
Model-based optimization of subgroup weights for survival analysis,
J. Richter, K. Madjar, and J. Rahnenführer, “Model-based optimization of subgroup weights for survival analysis,” Bioinformatics, vol. 35, no. 14, pp. i484–i491, 2019
2019
-
[98]
Loader, Local Regression and Likelihood
C. Loader, Local Regression and Likelihood. New Y ork, NY: Springer New Y ork, 1999, 290 pp
1999
-
[99]
Optimal weighted nearest neighbour classifiers,
R. J. Samworth, “Optimal weighted nearest neighbour classifiers,” Annals of Statistics, vol. 40, no. 5, pp. 2733–2763, 2012
2012
-
[100]
Assessing nonsuperiority, noninferiority, or equivalence when comparing two regression models over a restricted covariate region,
W. Liu, F . Bretz, A. J. Hayter, and H. P . Wynn, “Assessing nonsuperiority, noninferiority, or equivalence when comparing two regression models over a restricted covariate region,”Bio- metrics, vol. 65, no. 4, pp. 1279–1287, 2009
2009
-
[101]
Equivalence of regression curves,
H. Dette, K. Möllenhoff, S. Volgushev, and F . Bretz, “Equivalence of regression curves,”Jour- nal of the American Statistical Association, vol. 113, no. 522, pp. 711–729, 2018
2018
-
[102]
Similarity of competing risks models with constant intensities in an application to clinical healthcare pathways involving prostate cancer surgery,
N. Binder, K. Möllenhoff, A. Sigle, and H. Dette, “Similarity of competing risks models with constant intensities in an application to clinical healthcare pathways involving prostate cancer surgery,” Statistics in Medicine, vol. 41, no. 19, pp. 3804–3819, 2022. 32
2022
-
[103]
Data mining in urology: Understanding real-world treatment pathways for lower urinary tract systems via exploration of big data,
N. Binder, H. Dette, J. Franz, D. Zöller, R. Suarez-Ibarrola, C. Gratzke, et al., “Data mining in urology: Understanding real-world treatment pathways for lower urinary tract systems via exploration of big data,”European Urology Focus, vol. 8, no. 2, pp. 391–393, 2022
2022
-
[104]
Risk prediction models: II. external validation, model updating, and impact assessment,
K. G. M. Moons, A. P . Kengne, D. E. Grobbee, P . Royston, Y . Vergouwe, D. G. Altman, et al., “Risk prediction models: II. external validation, model updating, and impact assessment,” Heart, vol. 98, no. 9, pp. 691–698, 2012
2012
-
[105]
Developing more generalizable prediction models from pooled studies and large clustered data sets,
V. M. T. de Jong, K. G. M. Moons, M. J. C. Eijkemans, R. D. Riley, and T. P . A. Debray, “Developing more generalizable prediction models from pooled studies and large clustered data sets,” Statistics in Medicine, vol. 40, no. 15, pp. 3533–3559, 2021
2021
-
[106]
Characterizing patterns of care using administrative claims data: ADHD treatment in children,
G. R. Klein, J. B. Greenhouse, B. D. Stein, and H. J. Seltman, “Characterizing patterns of care using administrative claims data: ADHD treatment in children,” Health Services and Outcomes Research Methodology, vol. 11, no. 3, pp. 115–133, 2011
2011
-
[107]
Similarity measure between patient traces for clin- ical pathway analysis: Problem, method, and applications,
Z. Huang, W. Dong, H. Duan, and H. Li, “Similarity measure between patient traces for clin- ical pathway analysis: Problem, method, and applications,” IEEE Journal of Biomedical and Health Informatics, vol. 18, no. 1, pp. 4–14, 2014
2014
-
[108]
Visual semantic segmentation based on few/zero-shot learning: An overview,
W. Ren, Y . Tang, Q. Sun, C. Zhao, and Q.-L. Han, “Visual semantic segmentation based on few/zero-shot learning: An overview,” IEEE/CAA Journal of Automatica Sinica, pp. 1–21, 2023
2023
-
[109]
Learning similarity measures from data,
B. M. Mathisen, A. Aamodt, K. Bach, and H. Langseth, “Learning similarity measures from data,” Progress in Artificial Intelligence, vol. 9, no. 2, pp. 129–143, 2020
2020
-
[110]
Learning a hybrid similarity measure for image retrieval,
J. Wu, H. Shen, Y .-D. Li, Z.-B. Xiao, M.-Y . Lu, and C.-L. Wang, “Learning a hybrid similarity measure for image retrieval,”Pattern Recognition, vol. 46, no. 11, pp. 2927–2939, 2013
2013
-
[111]
Similarity learning for high-dimensional sparse data,
K. Liu, A. Bellet, and F . Sha, “Similarity learning for high-dimensional sparse data,” in In- ternational Conference on Artificial Intelligence and Statistics , 2015, pp. 653–662. [Online]. Available: https://proceedings.mlr.press/v38/liu15.html
2015
-
[112]
sc- VAE: Variational auto-encoders for single-cell gene expression data,
C. H. Grønbech, M. F . Vording, P . Timshel, C. K. Sønderby, T. H. Pers, and O. Winther, “sc- VAE: Variational auto-encoders for single-cell gene expression data,”Bioinformatics, vol. 36, no. 16, pp. 4415–4422, 2020
2020
-
[113]
Deep dynamic modeling with just two time points: Can we still allow for individual trajecto- ries?
M. Hackenberg, P . Harms, M. Pfaffenlehner, A. Pechmann, J. Kirschner, T. Schmidt, et al., “Deep dynamic modeling with just two time points: Can we still allow for individual trajecto- ries?” Biometrical Journal, 2022
2022
-
[114]
Dataset2Vec: Learning dataset meta- features,
H. S. Jomaa, L. Schmidt-Thieme, and J. Grabocka, “Dataset2Vec: Learning dataset meta- features,” Data Mining and Knowledge Discovery, vol. 35, no. 3, pp. 964–985, 2021
2021
-
[115]
Planning future studies based on the preci- sion of network meta-analysis results,
A. Nikolakopoulou, D. Mavridis, and G. Salanti, “Planning future studies based on the preci- sion of network meta-analysis results,” Statistics in Medicine, vol. 35, no. 7, pp. 978–1000, 2016
2016
-
[116]
An easy and efficient approach for testing identifiability,
C. Kreutz, “An easy and efficient approach for testing identifiability,” Bioinformatics, vol. 34, no. 11, pp. 1913–1921, 2018. 33
1913
-
[117]
Neural architecture evolution in deep reinforcement learning for continuous control,
J. K. H. Franke, G. Köhler, N. Awad, and F . Hutter, “Neural architecture evolution in deep reinforcement learning for continuous control,” 2020. arXiv: 1910.12824, pre-published
2020 arXiv
-
[118]
Hyperparameter ensembles for robust- ness and uncertainty quantification,
F . Wenzel, J. Snoek, D. Tran, and R. Jenatton, “Hyperparameter ensembles for robust- ness and uncertainty quantification,” inAdvances in Neural Information Processing Systems, 2020, pp. 6514–6527. [Online]. Available: https://papers.neurips.cc/paper_files/pa per/2020/hash/481...
2020
-
[119]
Hyperparameter op- timization: Foundations, algorithms, best practices, and open challenges,
B. Bischl, M. Binder, M. Lang, T. Pielok, J. Richter, S. Coors, et al., “Hyperparameter op- timization: Foundations, algorithms, best practices, and open challenges,” Wiley Interdisci- plinary Reviews: Data Mining and Knowledge Discovery, vol. 13, no. 2, e1484, 2023
2023
-
[120]
On the power of foundation models,
Y . Yuan, “On the power of foundation models,” inInternational Conference on Machine Learn- ing, 2023, pp. 40 519–40 530. [Online]. Available:https://proceedings.mlr.press/v202 /yuan23b.html
2023
-
[121]
Using self-supervised learning can improve model robustness and uncertainty,
D. Hendrycks, M. Mazeika, S. Kadavath, and D. Song, “Using self-supervised learning can improve model robustness and uncertainty,” in Advances in Neural Information Processing Systems, 2019. [Online]. Available: https://proceedings.neurips.cc/paper_files/pa per/2019/file/a2b15...
2019
-
[122]
Scientific discovery in the age of artificial intelligence,
H. Wang, T. Fu, Y . Du, W. Gao, K. Huang, Z. Liu, et al., “Scientific discovery in the age of artificial intelligence,” Nature, vol. 620, no. 7972, pp. 47–60, 2023
2023
-
[123]
Attention is all you need,
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, et al., “Attention is all you need,” in Advances in Neural Information Processing Systems , 2017. [Online]. Available: https://proceedings.neurips.cc/paper_files/paper/2017/hash/3f5ee24 3547dee91fbd053c1c4...
2017
-
[124]
Blockwise parallel decoding for deep autoregressive models,
M. Stern, N. Shazeer, and J. Uszkoreit, “Blockwise parallel decoding for deep autoregressive models,” in Advances in Neural Information Processing Systems , 2018. [Online]. Available: https://proceedings.neurips.cc/paper_files/paper/2018/hash/c4127b9194fe856 2c64dc0f5bf2c93bc-...
2018
-
[125]
Improved transformer for high- resolution gans,
L. Zhao, Z. Zhang, T. Chen, D. Metaxas, and H. Zhang, “Improved transformer for high- resolution gans,” in Advances in Neural Information Processing Systems, 2021, pp. 18 367– 18 380. [Online]. Available: https://proceedings.neurips.cc/paper/2021/hash/98dce 83da57b0395e163467c...
2021
-
[126]
TransGAN: Two pure transformers can make one strong GAN, and that can scale up,
Y . Jiang, S. Chang, and Z. Wang, “TransGAN: Two pure transformers can make one strong GAN, and that can scale up,” in Advances in Neural Information Processing Systems, 2021, pp. 14 745–14 758. [Online]. Available:https://proceedings.neurips.cc/paper/2021/h ash/7c220a2091c26a...
2021
-
[127]
ST -GAN: Spatial transformer generative adversarial networks for image compositing,
C.-H. Lin, E. Yumer, O. Wang, E. Shechtman, and S. Lucey, “ST -GAN: Spatial transformer generative adversarial networks for image compositing,” in IEEE/CVF Conference on Com- puter Vision and Pattern Recognition, 2018, pp. 9455–9464. 34
2018
-
[128]
TTS-GAN: A transformer-based time-series generative adversarial network,
X. Li, V. Metsis, H. Wang, and A. H. H. Ngu, “TTS-GAN: A transformer-based time-series generative adversarial network,” in Artificial Intelligence in Medicine , Springer International Publishing, 2022, pp. 133–143
2022
-
[129]
Language models are unsupervised multitask learners,
A. Radford, J. Wu, R. Child, D. Luan, D. Amodei, and I. Sutskever, “Language models are unsupervised multitask learners,” OpenAI Blog, 2019. [Online]. Available: https://cdn.op enai.com/better-language-models/language_models_are_unsupervised_multitask _learners.pdf
2019
-
[130]
A brief overview of ChatGPT: The his- tory, status quo and potential future development,
T. Wu, S. He, J. Liu, S. Sun, K. Liu, Q.-L. Han, et al., “A brief overview of ChatGPT: The his- tory, status quo and potential future development,” IEEE/CAA Journal of Automatica Sinica, vol. 10, no. 5, pp. 1122–1136, 2023
2023
-
[131]
BERT: Pre-training of deep bidirec- tional transformers for language understanding,
J. Devlin, M.-W. Chang, K. Lee, and K. Toutanova, “BERT: Pre-training of deep bidirec- tional transformers for language understanding,” inConference of the North American Chap- ter of the Association for Computational Linguistics: Human Language Technologies , 2019, pp. 4171–4186
2019
-
[132]
Single- sequence protein structure prediction using a language model and deep learning,
R. Chowdhury, N. Bouatta, S. Biswas, C. Floristean, A. Kharkar, K. Roy, et al. , “Single- sequence protein structure prediction using a language model and deep learning,” Nature Biotechnology, vol. 40, no. 11, pp. 1617–1623, 2022
2022
-
[133]
A survey on vision transformer,
K. Han, Y . Wang, H. Chen, X. Chen, J. Guo, Z. Liu, et al., “A survey on vision transformer,” IEEE Transactions on Pattern Analysis and Machine Intelligence, vol. 45, no. 1, pp. 87–110, 2022
2022
-
[134]
Transformers in vision: A survey,
S. Khan, M. Naseer, M. Hayat, S. W. Zamir, F . S. Khan, and M. Shah, “Transformers in vision: A survey,” ACM Computing Surveys (CSUR), vol. 54, pp. 1–41, 10s 2022
2022
-
[135]
Transformers can do Bayesian inference,
S. Müller, N. Hollmann, S. P . Arango, J. Grabocka, and F . Hutter, “Transformers can do Bayesian inference,” inInternational Conference on Learning Representation, 2022. [Online]. Available: https://openreview.net/pdf?id=KSugKcbNf9
2022
-
[136]
TabPFN: A transformer that solves small tabular classification problems in a second,
N. Hollmann, S. Müller, K. Eggensperger, and F . Hutter, “TabPFN: A transformer that solves small tabular classification problems in a second,” in International Conference on Learning Representations, 2022. [Online]. Available: https://openreview.net/forum?id=cp5Pvc I6w8_
2022
-
[137]
Accu- rate predictions on small data with a tabular foundation model,
N. Hollmann, S. Müller, L. Purucker, A. Krishnakumar, M. Körfer, S. B. Hoo, et al., “Accu- rate predictions on small data with a tabular foundation model,” Nature, vol. 637, no. 8045, pp. 319–326, 2025
2025
-
[138]
Retrieval-augmented generation for knowledge-intensive NLP tasks,
P . Lewis, E. Perez, A. Piktus, F . Petroni, V. Karpukhin, N. Goyal,et al., “Retrieval-augmented generation for knowledge-intensive NLP tasks,” in Advances in Neural Information Process- ing Systems, 2020, pp. 9459–9474. [Online]. Available: https://proceedings.neurips.c c/pap...
2020
-
[139]
Retrieval-augmented generation for large language models: A survey,
Y . Gao, Y . Xiong, X. Gao, K. Jia, J. Pan, Y . Bi, et al., “Retrieval-augmented generation for large language models: A survey,” 2024. arXiv: 2312.10997, pre-published. 35
2024 arXiv
-
[140]
Benchmarking large language models in retrieval- augmented generation,
J. Chen, H. Lin, X. Han, and L. Sun, “Benchmarking large language models in retrieval- augmented generation,” Conference on Artificial Intelligence , vol. 38, no. 16, pp. 17 754– 17 762, 2024
2024
-
[141]
Few-shot continual learning: A brain-inspired ap- proach,
L. Wang, Q. Li, Y . Zhong, and J. Zhu, “Few-shot continual learning: A brain-inspired ap- proach,” 2021. arXiv: 2104.09034, pre-published
2021 arXiv
-
[142]
Research progress on few-shot learning for remote sensing image interpretation,
X. Sun, B. Wang, Z. Wang, H. Li, H. Li, and K. Fu, “Research progress on few-shot learning for remote sensing image interpretation,” IEEE Journal of Selected Topics in Applied Earth Observations and Remote Sensing, vol. 14, pp. 2387–2402, 2021
2021
-
[143]
ImageNet: A large-scale hierar- chical image database,
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei, “ImageNet: A large-scale hierar- chical image database,” in IEEE Conference on Computer Vision and Pattern Recognition , IEEE, 2009, pp. 248–255
2009
-
[144]
ChestX-ray8: Hospital-scale chest x-ray database and benchmarks on weakly-supervised classification and localization of common thorax diseases,
X. Wang, Y . Peng, L. Lu, Z. Lu, M. Bagheri, and R. M. Summers, “ChestX-ray8: Hospital-scale chest x-ray database and benchmarks on weakly-supervised classification and localization of common thorax diseases,” inIEEE Conference on Computer Vision and Pattern Recognition, IEEE, 2017
2017
-
[145]
FHIST: A benchmark for few-shot classification of histological images,
F . Shakeri, M. Boudiaf, S. Mohammadi, I. Sheth, M. Havaei, I. B. Ayed, et al., “FHIST: A benchmark for few-shot classification of histological images,” 2022. arXiv: 2206.00092, pre- published
2022 arXiv
-
[146]
N. C. F . Codella, D. Gutman, M. E. Celebi, B. Helba, M. A. Marchetti, S. W. Dusza,et al., “Skin lesion analysis toward melanoma detection: A challenge at the 2017 international symposium on biomedical imaging (ISBI), hosted by the international skin imaging collaboration (ISI...
2017
-
[147]
The TCGA meta-dataset clinical benchmark,
M. Samiei, T. Würfl, T. Deleu, M. Weiss, F . Dutil, T. Fevens, et al., “The TCGA meta-dataset clinical benchmark,” 2019. arXiv: 1910.08636, pre-published
2019 arXiv
-
[148]
Model-agnostic meta-learning for fast adaptation of deep networks,
C. Finn, P . Abbeel, and S. Levine, “Model-agnostic meta-learning for fast adaptation of deep networks,” in International Conference on Machine Learning, 2017, pp. 1126–1135. [Online]. Available: http://proceedings.mlr.press/v70/finn17a.html
2017
-
[149]
Decoupled Weight Decay Regularization,
I. Loshchilov and F . Hutter, “Decoupled Weight Decay Regularization,” presented at the In- ternational Conference on Learning Representations, 2018. [Online]. Available: https://op enreview.net/forum?id=Bkg6RiCqY7
2018
-
[150]
Meta-learning,
J. Vanschoren, “Meta-learning,” in Automated Machine Learning: Methods, Systems, Chal- lenges, Cham: Springer International Publishing, 2019, pp. 35–61
2019
-
[151]
An overview of deep learning architectures in few-shot learning domain,
S. Jadon and A. Jadon, “An overview of deep learning architectures in few-shot learning domain,” 2023. arXiv: 2008.06365, pre-published
2023 arXiv
-
[152]
Hitzler and M
P . Hitzler and M. K. Sarker, Neuro-Symbolic Artificial Intelligence: The State of the Art. Ams- terdam Berlin Washington, DC: IOS Press, 2022, 408 pp
2022
-
[153]
Neuro-symbolic approaches in artificial intelligence,
P . Hitzler, A. Eberhart, M. Ebrahimi, M. K. Sarker, and L. Zhou, “Neuro-symbolic approaches in artificial intelligence,” National Science Review, vol. 9, no. 6, nwac035, 2022. 36
2022
-
[154]
Neurosymbolic AI: The 3rd wave,
A. d’Avila Garcez and L. C. Lamb, “Neurosymbolic AI: The 3rd wave,” Artificial Intelligence Review, vol. 56, no. 11, pp. 12 387–12 406, 2023
2023
-
[155]
Neurosymbolic artificial intelligence (why, what, and how),
A. Sheth, K. Roy, and M. Gaur, “Neurosymbolic artificial intelligence (why, what, and how),” IEEE Intelligent Systems, vol. 38, no. 03, pp. 56–62, 2023
2023
-
[156]
The nature of systems biology,
F . J. Bruggeman and H. V. Westerhoff, “The nature of systems biology,” Trends in Microbiol- ogy, vol. 15, no. 1, pp. 45–50, 2007
2007
-
[157]
Machine learning meets omics: Applications and perspec- tives,
R. Li, L. Li, Y . Xu, and J. Y ang, “Machine learning meets omics: Applications and perspec- tives,”Briefings in Bioinformatics, vol. 23, no. 1, pp. 1–22, 2022
2022
-
[158]
Neural ordinary differential equations,
R. T. Q. Chen, Y . Rubanova, J. Bettencourt, and D. Duvenaud, “Neural ordinary differential equations,” 2019. arXiv: 1806.07366, pre-published
2019 arXiv
-
[159]
Individualizing deep dynamic models for psychological resilience data,
G. Köber, S. Pooseh, H. Engen, A. Chmitorz, M. Kampa, A. Schick, et al., “Individualizing deep dynamic models for psychological resilience data,” Scientific Reports , vol. 12, no. 1, p. 8061, 2022
2022
-
[160]
Universal dif- ferential equations for scientific machine learning,
C. Rackauckas, Y . Ma, J. Martensen, C. Warner, K. Zubov, R. Supekar, et al., “Universal dif- ferential equations for scientific machine learning,” 2021. arXiv: 2001.04385, pre-published
2021 arXiv
-
[161]
Biologically- informed neural networks guide mechanistic modeling from sparse experimental data,
J. H. Lagergren, J. T. Nardini, R. E. Baker, M. J. Simpson, and K. B. Flores, “Biologically- informed neural networks guide mechanistic modeling from sparse experimental data,”PLOS Computational Biology, vol. 16, no. 12, e1008462, 2020
2020
-
[162]
A differentiable programming system to bridge machine learning and scientific computing,
M. Innes, A. Edelman, K. Fischer, C. Rackauckas, E. Saba, V. B. Shah,et al., “A differentiable programming system to bridge machine learning and scientific computing,” 2019. arXiv: 190 7.07587, pre-published
2019
-
[163]
Bayesian statistics and modelling,
R. van de Schoot, S. Depaoli, R. King, B. Kramer, K. Märtens, M. G. Tadesse,et al., “Bayesian statistics and modelling,” Nature Reviews Methods Primers, vol. 1, no. 1, pp. 1–26, 2021
2021
-
[164]
Blending diverse physical priors with neural networks,
Y . Ba, G. Zhao, and A. Kadambi, “Blending diverse physical priors with neural networks,”
-
[165]
Applications of deep learning and domain adaptation to the study of mountain snow,
J. A. Gongora, H. P . Marshall, C. Hase, T. Bradberry, and S. Ballerini, “Applications of deep learning and domain adaptation to the study of mountain snow,” American Geophysical Union, vol. 2019, 2019. [Online]. Available: https://ui.adsabs.harvard.edu/abs/20 19AGUFM.H33L2110G
2019
-
[166]
Deep hidden physics models: Deep learning of nonlinear partial differential equa- tions,
M. Raissi, “Deep hidden physics models: Deep learning of nonlinear partial differential equa- tions,” Journal of Machine Learning Research, vol. 19, p. 25, 2018. [Online]. Available: http s://www.jmlr.org/papers/volume19/18-046/18-046.pdf
2018
-
[167]
Informed pre- training on prior knowledge,
L. von Rueden, S. Houben, K. Cvejoski, C. Bauckhage, and N. Piatkowski, “Informed pre- training on prior knowledge,” 2022. arXiv: 2205.11433, pre-published
2022 arXiv
-
[168]
Informed machine learning – a taxonomy and survey of integrating prior knowledge into learning sys- tems,
L. von Rueden, S. Mayer, K. Beckh, B. Georgiev, S. Giesselbach, R. Heese, et al., “Informed machine learning – a taxonomy and survey of integrating prior knowledge into learning sys- tems,” IEEE Transactions on Knowledge and Data Engineering, vol. 35, no. 1, pp. 614–633, 2023. 37
2023
-
[169]
Physics- informed machine learning,
G. E. Karniadakis, I. G. Kevrekidis, L. Lu, P . Perdikaris, S. Wang, and L. Y ang, “Physics- informed machine learning,” Nature Reviews Physics, vol. 3, no. 6, pp. 422–440, 2021
2021
-
[170]
Goodfellow, Y
I. Goodfellow, Y . Bengio, and A. Courville, Deep Learning. MIT Press, 2016. [Online]. Avail- able: http://www.deeplearningbook.org
2016
-
[171]
Addressing the class imbalance problem in medical datasets,
M. M. Rahman and D. N. Davis, “Addressing the class imbalance problem in medical datasets,” International Journal of Machine Learning and Computing, pp. 224–228, 2013
2013
-
[172]
Machine learning and artificial intelligence research for patient benefit: 20 critical questions on trans- parency, replicability, ethics, and effectiveness,
S. Vollmer, B. A. Mateen, G. Bohner, F . J. Király, R. Ghani, P . Jonsson, et al., “Machine learning and artificial intelligence research for patient benefit: 20 critical questions on trans- parency, replicability, ethics, and effectiveness,”BMJ, vol. 368, 2020
2020
-
[173]
The effect of training and testing process on machine learning in biomedical datasets,
M. K. Uçar, M. Nour, H. Sindi, K. Polat, et al., “The effect of training and testing process on machine learning in biomedical datasets,” Mathematical Problems in Engineering, vol. 2020, 2020
2020
-
[174]
Data leakage inflates prediction performance in connectome-based machine learning models,
M. Rosenblatt, L. Tejavibulya, R. Jiang, S. Noble, and D. Scheinost, “Data leakage inflates prediction performance in connectome-based machine learning models,” Nature Communi- cations, vol. 15, no. 1, p. 1829, 2024
2024
-
[175]
Machine learning: A new approach for dose individualization,
Q.-Y . Li, B.-H. Tang, Y .-E. Wu, B.-F . Y ao, W. Zhang, Y . Zheng,et al., “Machine learning: A new approach for dose individualization,” Clinical Pharmacology & Therapeutics, vol. 115, no. 4, pp. 727–744, 2024
2024
-
[176]
The FAIR guiding principles for scientific data management and stewardship,
M. D. Wilkinson, M. Dumontier, I. J. Aalbersberg, G. Appleton, M. Axton, A. Baak, et al., “The FAIR guiding principles for scientific data management and stewardship,” Scientific Data , vol. 3, no. 1, pp. 1–9, 2016. 38
2016
-
[2019]
arXiv: 1910.00201, pre-published
1910 arXiv
Reviewed August 6, 2026 · model on record in the stance chip above.
Discussion (0). Sign in to comment.