REVIEW 3 major objections 1 minor 1 cited by
Explicit near-optimal expanders make noisy k-XOR solvable in polynomial time by reducing it to Reed-Muller decoding.
Reviewed by Pith at T0; open to challenge. T0 means a machine referee read the full paper against a public rubric. the ladder, T0–T4 →
T0 review · grok-4.5
2026-07-13 11:02 UTC
load-bearing objection Abstract claims a clean counterexample to expansion-implies-hardness for noisy k-XOR via GUV+RM, but the packet contains the wrong full manuscript, so the reduction cannot be checked. the 3 major comments →
Expanders Meet Reed-Muller: Easy Instances of Noisy k-XOR
The pith
A machine-rendered reading of the paper's core claim, the machinery that carries it, and where it could break.
Core claim
There exist explicit families of k-left-regular bipartite graphs with near-optimal (N^{1-α},1-o(1))-expansion for which noisy k-XOR remains polynomial-time solvable at constant noise rate η=1/3 (with M=2^{O(log^{2} N)} and k=(log N)^{O(1)}). Under standard conjectures on Reed-Muller codes over the binary erasure channel the same holds for polynomial M and vanishing noise η=N^{-c}.
What carries the argument
Lossless expanders of Guruswami-Umans-Vadhan whose vertices are labeled so that the noisy XOR system becomes an instance of decoding Reed-Muller codes from random errors; known efficient RM decoders then solve the original problem.
Load-bearing premise
The particular algebraic labeling of the expander vertices must convert the noisy XOR instance into a genuine Reed-Muller decoding problem while preserving both the expansion parameters and the random-error model that existing decoders require.
What would settle it
Exhibit a concrete GUV-based graph family whose expansion meets the claimed (N^{1-α},1-o(1)) bound yet for which no polynomial-time RM decoder succeeds at noise rate 1/3, or prove that the labeling map fails to produce a Reed-Muller codeword plus random errors.
If this is right
- Expansion alone cannot be used as a hardness certificate for noisy k-XOR in Sum-of-Squares or low-degree polynomial models.
- Any proof of average-case hardness for noisy XOR must exploit structure beyond expansion of the constraint graph.
- The same GUV-plus-Reed-Muller template yields further easy instances once stronger RM decoding results become available.
- Concrete polynomial-time algorithms now exist for a non-trivial regime of parameters previously thought hard under expansion conjectures.
Where Pith is reading between the lines
- The construction suggests a broader program of “planting” efficient codes inside pseudorandom graphs to create counter-examples to expansion-based hardness conjectures for other CSPs.
- If the Reed-Muller BEC conjectures are confirmed, the same technique immediately upgrades the result to polynomial-size instances and vanishing noise, closing a larger portion of the conjectured hard regime.
- The reduction may also supply new average-case algorithms for related planted inference problems whose constraint graphs can be realized by the same expanders.
Editorial analysis
A structured set of objections, weighed in public.
Referee Report
Summary. The submission is presented as arXiv:2604.04188, claiming an explicit family of near-optimal expanders (based on Guruswami–Umans–Vadhan lossless expanders) for which noisy k-XOR is polynomial-time solvable by reducing, via an algebraic vertex labeling, to decoding Reed–Muller codes from random errors. Concretely it asserts poly-time algorithms at constant noise η=1/3 for M=2^{O(log² N)}, k=(log N)^{O(1)} and (N^{1-α},1-o(1))-expansion, with a conjectural extension to polynomial M and η=N^{-c}. The body of the manuscript supplied for review, however, is an entirely different paper (title “Uncertainty-Aware Foundation Models for Clinical Data,” different authors, arXiv stamp 2604.04175) on distributional latent representations for clinical multimodal data. No lemmas, constructions, expansion calculations, or runtime analyses for the claimed expander/RM reduction appear in the text.
Significance. If the abstract’s claims were substantiated, the result would be significant for average-case complexity and the theory of CSPs: it would separate expansion from hardness for noisy k-XOR and falsify natural conjectures linking SoS/low-degree lower bounds to expansion. The proposed combination of GUV expanders with 2010s RM random-error decoding is a natural and potentially powerful idea. Because the supplied full text is unrelated, none of these claims can be verified, and the significance of the actual submission cannot be assessed on its technical merits.
major comments (3)
- The full manuscript text does not correspond to the title, abstract, or arXiv identifier under review. The body is a clinical-ML paper on uncertainty-aware foundation models (different title, authors, and arXiv stamp 2604.04175). Consequently the central construction—algebraic labeling of GUV vertices that turns noisy k-XOR into RM decoding while preserving (N^{1-α},1-o(1))-expansion and the random-error noise model—cannot be checked. Without that reduction, the poly-time claims at η=1/3 (and the conjectural extension) are unsupported by any proof or algorithm in the document.
- All load-bearing quantitative claims in the abstract (M=2^{O(log² N)}, k=(log N)^{O(1)}, expansion (N^{1-α},1-o(1)), noise rate η=1/3, and the BEC-conjecture extension to M=N^{O(1)} and η=N^{-c}) lack any supporting lemmas, parameter calculations, or runtime analyses in the supplied text. Tables 1–6 and Sections 4–5 of the body address clinical AUROC/ECE metrics and have no relation to expanders or Reed–Muller codes.
- The abstract’s “key insight” (vertex interpretation of GUV graphs yielding an RM instance) is precisely the step that must be verified for correctness of the noise model and expansion parameters. That step is absent from the manuscript; the review therefore cannot confirm soundness of the claimed separation between expansion and hardness.
minor comments (1)
- The review packet appears to have concatenated the abstract of 2604.04188 with the full text of an unrelated clinical foundation-model paper (2604.04175). This packaging error alone prevents a technical review of the claimed result.
Circularity Check
No circularity: full text is an unrelated clinical-ML paper, so the claimed GUV-to-RM derivation cannot be audited; abstract alone shows a standard non-circular constructive reduction.
full rationale
The CACHEABLE PAPER SOURCE CONTEXT supplies the complete manuscript of an entirely different work (Uncertainty-Aware Foundation Models for Clinical Data, arXiv 2604.04175) whose authors, title, tables, and content have no relation to expanders, Reed-Muller codes, or noisy k-XOR. Consequently no lemmas, vertex labelings, expansion calculations, or decoder reductions from 2604.04188 are present to inspect. From the supplied abstract the claimed argument is an explicit construction: take the GUV lossless expanders, interpret their vertices so that the noisy k-XOR instance becomes an instance of decoding Reed-Muller codes from random errors, then invoke existing 2010s random-error RM decoders (or standard BEC conjectures). This is a combination of two independent prior bodies of work; it does not define a quantity in terms of the target claim, does not fit free parameters to data and then re-label the fit as a prediction, and does not rest on a self-citation of an unverified uniqueness theorem. Under the analyzer rules such external reliance is ordinary and non-circular. No reduction-by-construction can be exhibited, so the circularity score is 0 and the steps list is empty.
Axiom & Free-Parameter Ledger
free parameters (3)
- expansion exponent α
- noise rate η (1/3 or N^{-c})
- constraint density / degree exponents
axioms (3)
- domain assumption Lossless expanders of Guruswami-Umans-Vadhan exist with the stated expansion and degree parameters.
- domain assumption Reed-Muller codes can be decoded from large amounts of random errors (and, under conjecture, erasures) in polynomial time at the noise rates used.
- ad hoc to paper An appropriate algebraic interpretation of GUV vertices turns the noisy k-XOR instance into an RM decoding instance while preserving expansion and noise statistics.
read the original abstract
In the noisy $k$-XOR problem, one is given $y \in \mathbb{F}_2^M$ and must distinguish between $y$ uniform and $y = A x + e$, where $A$ is the adjacency matrix of a $k$-left-regular bipartite graph with $N$ variables and $M$ constraints, $x\in \mathbb{F}_2^N$ is random, and $e$ is noise with rate $\eta$. Lower bounds in restricted computational models such as Sum-of-Squares and low-degree polynomials are closely tied to the expansion of $A$, leading to conjectures that expansion implies hardness. We show that such conjectures are false by constructing an explicit family of graphs with near-optimal expansion for which noisy $k$-XOR is solvable in polynomial time. Our construction combines two powerful directions of work in pseudorandomness and coding theory that have not been previously put together. Specifically, our graphs are based on the lossless expanders of Guruswami, Umans and Vadhan (JACM 2009). Our key insight is that by an appropriate interpretation of the vertices of their graphs, the noisy XOR problem turns into the problem of decoding Reed-Muller codes from random errors. Then we build on a powerful body of work from the 2010s correcting from large amounts of random errors. Putting these together yields our construction. Concretely, we obtain explicit families for which noisy $k$-XOR is polynomial-time solvable at constant noise rate $\eta = 1/3$ for graphs with $M = 2^{O(\log^2 N)}$, $k = (\log N)^{O(1)}$, and $(N^{1-\alpha}, 1-o(1))$-expansion. Under standard conjectures on Reed-Muller codes over the binary erasure channel, this extends to families with $M = N^{O(1)}$, $k=(\log N)^{O(1)}$, expansion $(N^{1-\alpha}, 1-o(1))$ and polynomial-time algorithms at noise rate $\eta = N^{-c}$.
Forward citations
Cited by 1 Pith paper
-
The Polynomial-Time Low-Degree Conjecture is False
The polynomial-time low-degree conjecture is false: a permutation-invariant graph distribution with zero low-degree advantage through polylogarithmic degree can still be detected in polynomial time after edge resampling.
Reference graph
Works this paper leans on
-
[1]
Foundation model for advancing healthcare: challenges, opportunities and future directions.IEEE Reviews in Biomedical Engineering, 2024
Yuting He, Fuxiang Huang, Xinrui Jiang, Yuxiang Nie, Minghao Wang, Jiguang Wang, and Hao Chen. Foundation model for advancing healthcare: challenges, opportunities and future directions.IEEE Reviews in Biomedical Engineering, 2024
2024
-
[2]
Foundation models in bioinformatics.National science review, 12(4):nwaf028, 2025
Fei Guo, Renchu Guan, Yaohang Li, Qi Liu, Xiaowo Wang, Can Yang, and Jianxin Wang. Foundation models in bioinformatics.National science review, 12(4):nwaf028, 2025
2025
-
[3]
Foundation models defining a new era in vision: a survey and outlook.IEEE Transactions on Pattern Analysis and Machine Intelligence, 2025
Muhammad Awais, Muzammal Naseer, Salman Khan, Rao Muhammad Anwer, Hisham Cholakkal, Mubarak Shah, Ming-Hsuan Yang, and Fahad Shahbaz Khan. Foundation models defining a new era in vision: a survey and outlook.IEEE Transactions on Pattern Analysis and Machine Intelligence, 2025
2025
-
[4]
Foundation models for time series analysis: A tutorial and survey
Yuxuan Liang, Haomin Wen, Yuqi Nie, Yushan Jiang, Ming Jin, Dongjin Song, Shirui Pan, and Qingsong Wen. Foundation models for time series analysis: A tutorial and survey. InProceedings of the 30th ACM SIGKDD conference on knowledge discovery and data mining, pages 6555–6565, 2024
2024
-
[5]
A foundation model for intensive care: Unlocking generalization across tasks and domains at scale.medRxiv, pages 2025–07, 2025
Manuel Burger, Daphné Chopard, Malte Londschien, Fedor Sergeev, Hugo Yèche, Rita Kuznetsova, Martin Faltys, Eike Gerdes, Polina Leshetkina, Peter Bühlmann, et al. A foundation model for intensive care: Unlocking generalization across tasks and domains at scale.medRxiv, pages 2025–07, 2025
2025
-
[6]
Foundation models for time series forecasting.International IT Journal of Research, ISSN: 3007-6706, 2(4):144–156, 2024
Suresh Chandra Thakur. Foundation models for time series forecasting.International IT Journal of Research, ISSN: 3007-6706, 2(4):144–156, 2024
2024
-
[7]
A foundational vision transformer improves diagnostic performance for electrocardiograms.NPJ Digital Medicine, 6(1):108, 2023
Akhil Vaid, Joy Jiang, Ashwin Sawant, Stamatios Lerakis, Edgar Argulian, Yuri Ahuja, Joshua Lampert, Alexander Charney, Hayit Greenspan, Jagat Narula, et al. A foundational vision transformer improves diagnostic performance for electrocardiograms.NPJ Digital Medicine, 6(1):108, 2023
2023
-
[8]
Foundation models in healthcare: Opportunities, risks & strategies forward
Anja Thieme, Aditya Nori, Marzyeh Ghassemi, Rishi Bommasani, Tariq Osman Andersen, and Ewa Luger. Foundation models in healthcare: Opportunities, risks & strategies forward. InExtended abstracts of the 2023 CHI conference on human factors in computing systems, pages 1–4, 2023
2023
-
[9]
Foundation models for electronic health records: representation dynamics and transferability
Michael C Burkhart, Bashar Ramadan, Zewei Liao, Kaveri Chhikara, Juan C Rojas, William F Parker, and Brett K Beaulieu-Jones. Foundation models for electronic health records: representation dynamics and transferability. arXiv preprint arXiv:2504.10422, 2025
Pith/arXiv arXiv 2025
-
[10]
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. Bert: Pre-training of deep bidirectional transformers for language understanding. InProceedings of the 2019 conference of the North American chapter of the association for computational linguistics: human language technologies, volume 1 (long and short papers), pages 4171–4186, 2019
2019
-
[11]
Multi-scale 3d deep convolutional neural network for hyperspectral image classification
Mingyi He, Bo Li, and Huahui Chen. Multi-scale 3d deep convolutional neural network for hyperspectral image classification. In2017 IEEE International Conference on Image Processing (ICIP), pages 3904–3908. IEEE, 2017. 10
2017
-
[12]
Masked autoencoders are scalable vision learners
Kaiming He, Xinlei Chen, Saining Xie, Yanghao Li, Piotr Dollár, and Ross Girshick. Masked autoencoders are scalable vision learners. InProceedings of the IEEE/CVF conference on computer vision and pattern recognition, pages 16000–16009, 2022
2022
-
[13]
Deep residual learning for image recognition, 2015
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. Deep residual learning for image recognition, 2015
2015
-
[14]
Bag of tricks for image classification with convolutional neural networks
Tong He, Zhi Zhang, Hang Zhang, Zhongyue Zhang, Junyuan Xie, and Mu Li. Bag of tricks for image classification with convolutional neural networks. InProceedings of the IEEE/CVF conference on computer vision and pattern recognition, pages 558–567, 2019
2019
-
[15]
Cross attention network for few-shot classification.Advances in neural information processing systems, 32, 2019
Ruibing Hou, Hong Chang, Bingpeng Ma, Shiguang Shan, and Xilin Chen. Cross attention network for few-shot classification.Advances in neural information processing systems, 32, 2019
2019
-
[16]
Crossvit: Cross-attention multi-scale vision trans- former for image classification
Chun-Fu Richard Chen, Quanfu Fan, and Rameswar Panda. Crossvit: Cross-attention multi-scale vision trans- former for image classification. InProceedings of the IEEE/CVF international conference on computer vision, pages 357–366, 2021
2021
-
[17]
Ccnet: Criss-cross attention for semantic segmentation
Zilong Huang, Xinggang Wang, Lichao Huang, Chang Huang, Yunchao Wei, and Wenyu Liu. Ccnet: Criss-cross attention for semantic segmentation. InProceedings of the IEEE/CVF international conference on computer vision, pages 603–612, 2019
2019
-
[18]
Serialized ehr make for good text representations.arXiv preprint arXiv:2510.13843, 2025
Zhirong Chou, Quan Qin, and Shi Li. Serialized ehr make for good text representations.arXiv preprint arXiv:2510.13843, 2025
arXiv 2025
-
[19]
Clio: Policy-aware foundation models for ehr as controlled dynamical systems.Authorea Preprints, 2025
Fu Huiliang, Hu Hong, Tao Jingfei, Guo Fengge, Cai Ning, Yuanyun Zhang, and Shi Li. Clio: Policy-aware foundation models for ehr as controlled dynamical systems.Authorea Preprints, 2025
2025
-
[20]
Wu Hao Ran, Xi Xi, Furong Li, Jingyi Lu, Jian Jiang, Hui Huang, Yuzhuan Zhang, and Shi Li. Structured semantics from unstructured notes: Language model approaches to ehr-based decision support.arXiv preprint arXiv:2506.06340, 2025
Pith/arXiv arXiv 2025
-
[21]
Yuanyun Zhang and Shi Li. Chronoformer: Time-aware transformer architectures for structured clinical event modeling.arXiv preprint arXiv:2504.07373, 2025
Pith/arXiv arXiv 2025
-
[22]
Yuanyun Zhang and Shi Li. A collection of innovations in medical ai for patient records in 2024.arXiv preprint arXiv:2503.05768, 2025
Pith/arXiv arXiv 2024
-
[23]
Latent physiology as language: A state-space foundation model for multimodal icu and ehr representation learning
Shane Lowe, Garrett Park, Liam Lee, and Parker Smith. Latent physiology as language: A state-space foundation model for multimodal icu and ehr representation learning
-
[24]
Text as an inductive bias: A novel foundation model for electronic health records
Shi Li and Guang Dong. Text as an inductive bias: A novel foundation model for electronic health records. Authorea Preprints
-
[25]
Foundation models for physiological signals: Opportunities and challenges
Simon A Lee and Kai Akamatsu. Foundation models for physiological signals: Opportunities and challenges. August 2025
2025
-
[26]
Salar Abbaspourazad, Oussama Elachqar, Andrew C Miller, Saba Emrani, Udhyakumar Nallasamy, and Ian Shapiro. Large-scale training of foundation models for wearable biosignals.arXiv preprint arXiv:2312.05409, 2023
Pith/arXiv arXiv 2023
-
[27]
Gfmbench-api: A standardized interface for benchmarking genomic foundation models.bioRxiv, pages 2026–02, 2026
Ariel Larey, Elay Dahan, Amit Bleiweiss Amit Bleiweiss, Raizy Kellerman, Guy Leib, Omri Nayshool, Dan Ofer, Tal Zinger, Dan Dominissini, Gideon Rechavi, et al. Gfmbench-api: A standardized interface for benchmarking genomic foundation models.bioRxiv, pages 2026–02, 2026
2026
-
[28]
Mutbert: probabilistic genome representation improves genomics foundation models.bioinformatics, 41(Supplement_1):i294–i303, 2025
Weicai Long, Houcheng Su, Jiaqi Xiong, and Yanlin Zhang. Mutbert: probabilistic genome representation improves genomics foundation models.bioinformatics, 41(Supplement_1):i294–i303, 2025
2025
-
[29]
Sleepfm: Multi-modal representation learning for sleep across brain activity, ecg and respiratory signals
Rahul Thapa, Bryan He, Magnus Ruud Kjaer, Hyatt Moore Iv, Gauri Ganjoo, Emmanuel Mignot, and James Zou. Sleepfm: Multi-modal representation learning for sleep across brain activity, ecg and respiratory signals. In International Conference on Machine Learning, pages 48019–48037. PMLR, 2024
2024
-
[30]
Shovito Barua Soumma, Kartik Mangipudi, Daniel Peterson, Shyamal Mehta, and Hassan Ghasemzadeh. Wearable- based real-time freezing of gait detection in parkinson’s disease using self-supervised learning.arXiv preprint arXiv:2410.20715, 2024
Pith/arXiv arXiv 2024
-
[31]
Ariel Larey, Elay Dahan, Amit Bleiweiss, Raizy Kellerman, Guy Leib, Omri Nayshool, Dan Ofer, Tal Zinger, Dan Dominissini, Gideon Rechavi, et al. Jepa-dna: Grounding genomic foundation models through joint-embedding predictive architectures.arXiv preprint arXiv:2602.17162, 2026
Pith/arXiv arXiv 2026
-
[32]
Simon A Lee, Anthony Wu, and Jeffrey N Chiang. Clinical modernbert: An efficient and long context encoder for biomedical text.arXiv preprint arXiv:2504.03964, 2025. 11
Pith/arXiv arXiv 2025
-
[33]
Adibvafa Fallahpour, Mahshid Alinoori, Wenqian Ye, Xu Cao, Arash Afkanpour, and Amrit Krishnan. Ehrmamba: Towards generalizable and scalable foundation models for electronic health records.arXiv preprint arXiv:2405.14567, 2024
Pith/arXiv arXiv 2024
-
[34]
A simple framework for contrastive learning of visual representations
Ting Chen, Simon Kornblith, Mohammad Norouzi, and Geoffrey Hinton. A simple framework for contrastive learning of visual representations. InInternational conference on machine learning, pages 1597–1607. PMLR, 2020
2020
-
[35]
Contrastive representation distillation.arXiv, 2019
Yonglong Tian, Dilip Krishnan, and Phillip Isola. Contrastive representation distillation.arXiv, 2019
2019
-
[36]
Contrastive learning of preferences with a contextual infonce loss, 2024
Timo Bertram, Johannes Fürnkranz, and Martin Müller. Contrastive learning of preferences with a contextual infonce loss, 2024
2024
-
[37]
Clinical decision support using pseudo-notes from multiple streams of ehr data.npj Digital Medicine, 8(1):394, July 2025
Simon A Lee, Sujay Jain, Alex Chen, Kyoka Ono, Arabdha Biswas, Ákos Rudas, Jennifer Fang, and Jeffrey N Chiang. Clinical decision support using pseudo-notes from multiple streams of ehr data.npj Digital Medicine, 8(1):394, July 2025
2025
-
[38]
Med-bert: pretrained contextualized embeddings on large-scale structured electronic health records for disease prediction.NPJ digital medicine, 4(1):86, 2021
Laila Rasmy, Yang Xiang, Ziqian Xie, Cui Tao, and Degui Zhi. Med-bert: pretrained contextualized embeddings on large-scale structured electronic health records for disease prediction.NPJ digital medicine, 4(1):86, 2021
2021
-
[39]
Using foundation models to prescribe patients proper antibiotics
Simon A Lee, Helio Halperin, Yanai Halperin, Trevor Brokowski, and Jeffrey N Chiang. Using foundation models to prescribe patients proper antibiotics. InAAAI Bridge Program on AI for Medicine and Healthcare, pages 121–132. PMLR, 2025
2025
-
[40]
The shaky foundations of large language models and foundation models for electronic health records.npj digital medicine, 6(1):135, 2023
Michael Wornow, Yizhe Xu, Rahul Thapa, Birju Patel, Ethan Steinberg, Scott Fleming, Michael A Pfeffer, Jason Fries, and Nigam H Shah. The shaky foundations of large language models and foundation models for electronic health records.npj digital medicine, 6(1):135, 2023
2023
-
[41]
Simon A Lee, Sujay Jain, Alex Chen, Kyoka Ono, Jennifer Fang, Akos Rudas, and Jeffrey N Chiang. Emergency department decision support using clinical pseudo-notes.arXiv preprint arXiv:2402.00160, 2024
Pith/arXiv arXiv 2024
-
[42]
Language models are few-shot learners.Advances in neural information processing systems, 33:1877–1901, 2020
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Nee- lakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al. Language models are few-shot learners.Advances in neural information processing systems, 33:1877–1901, 2020
1901
-
[43]
Cehr-bert: Incorporating temporal information from structured ehr data to improve prediction tasks
Chao Pang, Xinzhuo Jiang, Krishna S Kalluri, Matthew Spotnitz, RuiJun Chen, Adler Perotte, and Karthik Natarajan. Cehr-bert: Incorporating temporal information from structured ehr data to improve prediction tasks. In Machine Learning for Health, pages 239–260. PMLR, 2021
2021
-
[44]
Chao Pang, Xinzhuo Jiang, Nishanth Parameshwar Pavinkurve, Krishna S Kalluri, Elise L Minto, Jason Patterson, Linying Zhang, George Hripcsak, Gamze Gürsoy, Noémie Elhadad, et al. Cehr-gpt: Generating electronic health records with chronological patient timelines.arXiv preprint arXiv:2402.04400, 2024
Pith/arXiv arXiv 2024
-
[45]
Event stream gpt: a data pre-processing and modeling library for generative, pre-trained transformers over continuous-time sequences of complex events
Matthew McDermott, Bret Nestor, Peniel Argaw, and Isaac S Kohane. Event stream gpt: a data pre-processing and modeling library for generative, pre-trained transformers over continuous-time sequences of complex events. Advances in Neural Information Processing Systems, 36, 2024
2024
-
[46]
Ummara Mumtaz, Awais Ahmed, and Summaya Mumtaz. Llms-healthcare: Current applications and challenges of large language models in various medical specialties.arXiv preprint arXiv:2311.12882, 2023
Pith/arXiv arXiv 2023
-
[47]
Llm4ts: Aligning pre-trained llms as data-efficient time-series forecasters.ACM Transactions on Intelligent Systems and Technology, 16(3):1–20, 2025
Ching Chang, Wei-Yao Wang, Wen-Chih Peng, and Tien-Fu Chen. Llm4ts: Aligning pre-trained llms as data-efficient time-series forecasters.ACM Transactions on Intelligent Systems and Technology, 16(3):1–20, 2025
2025
-
[48]
Accurate predictions on small data with a tabular foundation model.Nature, 637(8045):319–326, 2025
Noah Hollmann, Samuel Müller, Lennart Purucker, Arjun Krishnakumar, Max Körfer, Shi Bin Hoo, Robin Tibor Schirrmeister, and Frank Hutter. Accurate predictions on small data with a tabular foundation model.Nature, 637(8045):319–326, 2025
2025
-
[49]
Kyoka Ono and Simon A Lee. Text serialization and their relationship with the conventional paradigms of tabular machine learning.arXiv preprint arXiv:2406.13846, 2024
Pith/arXiv arXiv 2024
-
[50]
Clinical text summarization: Adapting large language models can outperform human experts.Research Square, 2023
Dave Van Veen, Cara Van Uden, Louis Blankemeier, Jean-Benoit Delbrouck, Asad Aali, Christian Bluethgen, Anuj Pareek, Malgorzata Polacin, Eduardo Pontes Reis, Anna Seehofnerova, et al. Clinical text summarization: Adapting large language models can outperform human experts.Research Square, 2023
2023
-
[51]
Time-llm: Time series forecasting by reprogramming large language models
Ming Jin, Shiyu Wang, Lintao Ma, Zhixuan Chu, James Y Zhang, Xiaoming Shi, Pin-Yu Chen, Yuxuan Liang, Yuan-Fang Li, Shirui Pan, et al. Time-llm: Time series forecasting by reprogramming large language models. arXiv preprint arXiv:2310.01728, 2023
Pith/arXiv arXiv 2023
-
[52]
Multimodal llms for health grounded in individual- specific data
Anastasiya Belyaeva, Justin Cosentino, Farhad Hormozdiari, Krish Eswaran, Shravya Shetty, Greg Corrado, Andrew Carroll, Cory Y McLean, and Nicholas A Furlotte. Multimodal llms for health grounded in individual- specific data. InWorkshop on Machine Learning for Multimodal Healthcare Data, pages 86–102. Springer, 2023. 12
2023
-
[53]
Yihan Lin, Zhirong Bella Yu, and Simon Lee. A case study exploring the current landscape of synthetic medical record generation with commercial llms.arXiv preprint arXiv:2504.14657, 2025
Pith/arXiv arXiv 2025
-
[54]
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. Deep residual learning for image recognition. In Proceedings of the IEEE conference on computer vision and pattern recognition, pages 770–778, 2016
2016
-
[55]
A computer-aided detection system for the detection of lung nodules based on 3d-resnet.Applied Sciences, 9(24):5544, 2019
Jiaxu Ning, Haitong Zhao, Lei Lan, Peng Sun, and Yunfei Feng. A computer-aided detection system for the detection of lung nodules based on 3d-resnet.Applied Sciences, 9(24):5544, 2019
2019
-
[56]
Introducing transfer learning to 3d resnet-18 for alzheimer’s disease detection on mri images
Amir Ebrahimi, Suhuai Luo, and Raymond Chiong. Introducing transfer learning to 3d resnet-18 for alzheimer’s disease detection on mri images. In2020 35th international conference on image and vision computing New Zealand (IVCNZ), pages 1–6. IEEE, 2020
2020
-
[57]
Automatic segmentation of head and neck (h&n) primary tumors in pet and ct images using 3d-inception-resnet model
Abdul Qayyum, Abdesslam Benzinou, Moona Mazher, Mohamed Abdel-Nasser, and Domenec Puig. Automatic segmentation of head and neck (h&n) primary tumors in pet and ct images using 3d-inception-resnet model. In 3D Head and Neck Tumor Segmentation in PET/CT Challenge, pages 58–67. Springer, 2021
2021
-
[58]
An image is worth 16x16 words: Transformers for image recognition at scale
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, Jakob Uszkoreit, and Neil Houlsby. An image is worth 16x16 words: Transformers for image recognition at scale. InInternational Conference on Learning Representations, 2021
2021
-
[59]
Swin transformer: Hierarchical vision transformer using shifted windows
Ze Liu, Yutong Lin, Yue Cao, Han Hu, Yixuan Wei, Zheng Zhang, Stephen Lin, and Baining Guo. Swin transformer: Hierarchical vision transformer using shifted windows. InProceedings of the IEEE/CVF international conference on computer vision, pages 10012–10022, 2021
2021
-
[60]
Swin unetr: Swin transformers for semantic segmentation of brain tumors in mri images
Ali Hatamizadeh, Vishwesh Nath, Yucheng Tang, Dong Yang, Holger R Roth, and Daguang Xu. Swin unetr: Swin transformers for semantic segmentation of brain tumors in mri images. InInternational MICCAI brainlesion workshop, pages 272–284. Springer, 2021
2021
-
[61]
Abdomenatlas: A large-scale, detailed-annotated, & multi-center dataset for efficient transfer learning and open algorithmic benchmarking.Medical Image Analysis, 97:103285, 2024
Wenxuan Li, Chongyu Qu, Xiaoxi Chen, Pedro RAS Bassi, Yijia Shi, Yuxiang Lai, Qian Yu, Huimin Xue, Yixiong Chen, Xiaorui Lin, et al. Abdomenatlas: A large-scale, detailed-annotated, & multi-center dataset for efficient transfer learning and open algorithmic benchmarking.Medical Image Analysis, 97:103285, 2024
2024
-
[62]
Guotai Wang, Jianghao Wu, Xiangde Luo, Xinglong Liu, Kang Li, and Shaoting Zhang. Mis-fm: 3d medical image segmentation using foundation models pretrained on a large-scale unannotated dataset.arXiv preprint arXiv:2306.16925, 2023
Pith/arXiv arXiv 2023
-
[63]
Linshan Wu, Jiaxin Zhuang, and Hao Chen. Large-scale 3d medical image pre-training with geometric context priors.arXiv preprint arXiv:2410.09890, 2024
Pith/arXiv arXiv 2024
-
[64]
Emerging properties in self-supervised vision transformers
Mathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou, Julien Mairal, Piotr Bojanowski, and Armand Joulin. Emerging properties in self-supervised vision transformers. InProceedings of the International Conference on Computer Vision (ICCV), 2021
2021
-
[65]
ibot: Image bert pre-training with online tokenizer.International Conference on Learning Representations (ICLR), 2022
Jinghao Zhou, Chen Wei, Huiyu Wang, Wei Shen, Cihang Xie, Alan Yuille, and Tao Kong. ibot: Image bert pre-training with online tokenizer.International Conference on Learning Representations (ICLR), 2022
2022
-
[66]
Dinov2: Learning robust visual features without supervision.arXiv preprint arXiv:2304.07193, 2023
Maxime Oquab, Timothée Darcet, Théo Moutakanni, Huy V o, Marc Szafraniec, Vasil Khalidov, Pierre Fernandez, Daniel Haziza, Francisco Massa, Alaaeldin El-Nouby, et al. Dinov2: Learning robust visual features without supervision.arXiv preprint arXiv:2304.07193, 2023
Pith/arXiv arXiv 2023
-
[67]
Learning transferable visual models from natural language supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al. Learning transferable visual models from natural language supervision. InInternational conference on machine learning, pages 8748–8763. PMLR, 2021
2021
-
[68]
FlashAttention-2: Faster attention with better parallelism and work partitioning
Tri Dao. FlashAttention-2: Faster attention with better parallelism and work partitioning. InInternational Conference on Learning Representations (ICLR), 2024
2024
-
[69]
Unetr++: delving into efficient and accurate 3d medical image segmentation.IEEE Transactions on Medical Imaging, 2024
Abdelrahman M Shaker, Muhammad Maaz, Hanoona Rasheed, Salman Khan, Ming-Hsuan Yang, and Fahad Shah- baz Khan. Unetr++: delving into efficient and accurate 3d medical image segmentation.IEEE Transactions on Medical Imaging, 2024
2024
-
[70]
Segmamba: Long-range sequential modeling mamba for 3d medical image segmentation
Zhaohu Xing, Tian Ye, Yijun Yang, Guang Liu, and Lei Zhu. Segmamba: Long-range sequential modeling mamba for 3d medical image segmentation. InInternational Conference on Medical Image Computing and Computer-Assisted Intervention, pages 578–588. Springer, 2024
2024
-
[71]
Zixuan Liu, Hanwen Xu, Addie Woicik, Linda G Shapiro, Marian Blazes, Yue Wu, Cecilia S Lee, Aaron Y Lee, and Sheng Wang. Octcube: a 3d foundation model for optical coherence tomography that improves cross-dataset, cross-disease, cross-device and cross-modality analysis.arXiv preprint arXiv:2408.11227, 2024. 13
Pith/arXiv arXiv 2024
-
[72]
Salar Abbaspourazad, Anshuman Mishra, Joseph Futoma, Andrew C Miller, and Ian Shapiro. Wearable accelerom- eter foundation models for health via knowledge distillation.arXiv preprint arXiv:2412.11276, 2024
Pith/arXiv arXiv 2024
-
[73]
Brandon Westover, and Jimeng Sun
Chaoqi Yang, M. Brandon Westover, and Jimeng Sun. Biot: Biosignal transformer for cross-data learning in the wild. InNeurIPS 2023, 2023
2023
-
[74]
Pearson Education India, 1999
Alan V Oppenheim.Discrete-time signal processing. Pearson Education India, 1999
1999
-
[75]
SIAM, 1992
Ingrid Daubechies.Ten lectures on wavelets. SIAM, 1992
1992
-
[76]
Towards on-device foundation models for raw wearable signals
Simon A Lee, Cyrus Tanade, Hao Zhou, Juhyeon Lee, Megha Thukral, Baiying Lu, and Sharanya Arcot Desai. Towards on-device foundation models for raw wearable signals. InNeurIPS 2025 Workshop on Learning from Time Series for Health, 2025
2025
-
[77]
Simon A Lee, Cyrus Tanade, Hao Zhou, Juhyeon Lee, Megha Thukral, Minji Han, Rachel Choi, Md Sazzad Hissain Khan, Baiying Lu, Migyeong Gwak, et al. Himae: Hierarchical masked autoencoders discover resolution-specific structure in wearable time series.arXiv preprint arXiv:2510.25785, 2025
arXiv 2025
-
[78]
Meds: Building models and tools in a reproducible health ai ecosystem
Matthew BA McDermott, Justin Xu, Teya S Bergamaschi, Hyewon Jeong, Simon A Lee, Nassim Oufattole, Patrick Rockenschaub, Kamil˙e Stankeviˇci¯ut˙e, Ethan Steinberg, Jimeng Sun, et al. Meds: Building models and tools in a reproducible health ai ecosystem. InProceedings of the 31st ACM SIGKDD Conference on Knowledge Discovery and Data Mining V . 2, pages 6243...
2025
-
[79]
Meds decentralized, extensible validation (meds-dev) benchmark: Establishing reproducibility and comparability in ml for health
Aleksia Kolo, Chao Pang, Edward Choi, Ethan Steinberg, Hyewon Jeong, Jack Gallifant, Jason A Fries, Jeffrey N Chiang, Jungwoo Oh, Justin Xu, et al. Meds decentralized, extensible validation (meds-dev) benchmark: Establishing reproducibility and comparability in ml for health. 2024
2024
-
[80]
Michael Wornow, Suhana Bedi, Miguel Angel Fuentes Hernandez, Ethan Steinberg, Jason Alan Fries, Christopher Ré, Sanmi Koyejo, and Nigam H Shah. Context clues: Evaluating long context models for clinical prediction tasks on ehrs.arXiv preprint arXiv:2412.16178, 2024
Pith/arXiv arXiv 2024
discussion (0)
Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.