REVIEW 1 major objections 1 minor 2 cited by
Joint dynamic programming co-designs sensor geometry and adaptive policy to exceed non-adaptive information limits.
Reviewed by Pith at T0; open to challenge. T0 means a machine referee read the full paper against a public rubric. the ladder, T0–T4 →
T0 review · grok-4.3
2026-07-01 09:15 UTC pith:HWHWXDDG
load-bearing objection The paper's joint-DP formulation for co-optimizing sensor geometry and adaptive policy is new, but the central claim rests on unverified gradient accuracy through the Bellman max at scale. the 1 major comments →
Adaptive Sensing beyond Non-Adaptive Information Limits: End-to-End Co-Design of Geometry, Policy, and Inference
The pith
A machine-rendered reading of the paper's core claim, the machinery that carries it, and where it could break.
Core claim
The authors claim that joint dynamic programming over sensor geometry and a Bellman-optimal adaptive measurement policy enables adaptive sensing to surpass the information limits of any non-adaptive strategy, with the outer hardware gradient obtained through differentiable dynamic programming that employs a sharp Bellman maximum; a hierarchy of relaxations then extends the same framework from small POMDPs to freeform photonic topologies with more than 10^5 design pixels.
What carries the argument
joint dynamic programming (joint-DP), a unified optimization that treats sensor geometry and the adaptive policy as a single dynamic program whose outer gradient is computed via differentiable dynamic programming with a sharp Bellman maximum
Load-bearing premise
The outer hardware gradient obtained through differentiable dynamic programming with a sharp Bellman maximum remains accurate and stable when applied to the joint optimization of geometry and policy.
What would settle it
A numerical experiment in which the jointly optimized geometry-plus-policy system fails to capture more information than either an optimized fixed geometry with a separately trained policy or a fixed geometry with an optimized adaptive policy, or in which the computed hardware gradients diverge from finite-difference checks.
If this is right
- Adaptive sensing on jointly designed hardware can exceed the information capture achievable by non-adaptive strategies on the same hardware.
- The same optimization framework scales, via successive relaxations, to freeform photonic devices containing more than 10^5 design pixels.
- Intelligence previously located in downstream digital algorithms can be embedded directly in the physical sensing structure.
- The hierarchy of relaxations provides a systematic path from small discrete decision problems to large continuous photonic design tasks.
Where Pith is reading between the lines
- The same joint-optimization logic could be applied to acoustic or RF sensing modalities by replacing the photonic wave solver with the appropriate physics model.
- Physical structures might be viewed as carrying an embedded policy that reduces the computational load on any attached digital processor.
- Experimental validation would require fabricating the jointly optimized geometry and measuring actual information gain against sequentially optimized baselines.
- The approach suggests end-to-end pipelines in which the physical layer and the inference algorithm are trained together rather than in stages.
Editorial analysis
A structured set of objections, weighed in public.
Referee Report
Summary. The paper claims that joint dynamic programming (joint-DP) enables co-design of sensor geometry and a Bellman-optimal adaptive measurement policy, allowing adaptive sensing to exceed non-adaptive information limits. The outer hardware gradient is obtained via differentiable dynamic programming using a sharp Bellman maximum, and a hierarchy of relaxations scales the approach to freeform photonic designs exceeding 10^5 pixels.
Significance. If the claimed gradients remain accurate and stable under joint optimization, the framework would represent a meaningful advance by relocating adaptive intelligence into the physical layer of sensors, with direct implications for information-limited sensing tasks. The hierarchy of relaxations is a notable technical contribution for scaling to large design spaces, though no machine-checked proofs or reproducible code are provided to substantiate the claims.
major comments (1)
- [Abstract] Abstract (and method description): The central claim relies on obtaining an accurate outer hardware gradient by differentiating through dynamic programming that employs a sharp Bellman maximum. No surrogate for the non-differentiable max operator is specified, nor is there a proof or bound showing that the resulting gradient remains unbiased or that approximation error does not grow with state-space size or across the hierarchy of relaxations needed for >10^5-pixel topologies. This directly impacts the validity of the joint optimization.
minor comments (1)
- The abstract states the method extends from small discrete POMDPs to freeform topologies, but provides no concrete example or scaling plot to illustrate the hierarchy of relaxations.
Simulated Author's Rebuttal
We thank the referee for their careful review and constructive feedback. We address the single major comment point-by-point below and will revise the manuscript accordingly to improve clarity on the gradient computation.
read point-by-point responses
-
Referee: [Abstract] Abstract (and method description): The central claim relies on obtaining an accurate outer hardware gradient by differentiating through dynamic programming that employs a sharp Bellman maximum. No surrogate for the non-differentiable max operator is specified, nor is there a proof or bound showing that the resulting gradient remains unbiased or that approximation error does not grow with state-space size or across the hierarchy of relaxations needed for >10^5-pixel topologies. This directly impacts the validity of the joint optimization.
Authors: We agree that the abstract is concise and does not detail the implementation. In the body of the manuscript the sharp Bellman maximum is realized by exact argmax selection over actions, with gradients back-propagated only through the optimal action (standard subgradient handling in autodiff frameworks, with zero gradient on non-optimal branches). No softmax-style surrogate is employed in order to preserve sharpness. We acknowledge that explicit theoretical bounds on bias or error growth with state-space size or across the relaxation hierarchy are not derived. We will revise the manuscript to add a dedicated paragraph (or short appendix) clarifying the argmax gradient flow, describing how the hierarchy of relaxations preserves differentiability, and reporting empirical checks of gradient stability on designs up to 10^5 pixels. This revision will directly address the validity concern without altering the claimed results. revision: yes
Circularity Check
No circularity: joint-DP is a new optimization construction
full rationale
The provided abstract and description present joint dynamic programming as a novel unified optimization over geometry and Bellman-optimal policy, with the outer gradient obtained via differentiable DP. No equations or steps are shown that reduce by construction to fitted inputs, self-citations, or renamed known results. The hierarchy of relaxations is framed as an extension to larger topologies rather than a self-referential derivation. The central claim remains an independent methodological proposal without load-bearing reductions to its own inputs.
Axiom & Free-Parameter Ledger
axioms (1)
- domain assumption Differentiable dynamic programming with a sharp Bellman maximum yields usable gradients for joint hardware-policy optimization.
Cite this review
Pith. "Pith review of Adaptive Sensing beyond Non-Adaptive Information Limits: End-to-End Co-Design of Geometry, Policy, and Inference." pith.science (2026). https://pith.science/paper/HWHWXDDG
@misc{pith2026260425193,
author = {Pith},
title = {Pith review of: Adaptive Sensing beyond Non-Adaptive Information Limits: End-to-End Co-Design of Geometry, Policy, and Inference},
year = {2026},
howpublished = {\url{https://pith.science/paper/HWHWXDDG}},
note = {Machine review of arXiv:2604.25193}
}
read the original abstract
Inverse design has transformed vast physical parameter spaces into a substrate for emergent functionality, raising the tantalizing prospect of relocating intelligence from the digital domain into the physical world itself. Nowhere is this prospect more consequential than in sensing, where the analog-to-digital interface imposes a fundamental bottleneck: information not captured by the hardware is irrevocably lost to any downstream algorithm. Existing approaches improve information capture through either sensor hardware optimization or adaptive measurement strategies operating on fixed hardware, but rarely both in concert. A principled migration of intelligence from digital to physical demands their joint optimization: the sensing geometry must be co-designed with a policy that determines what to measure next. We formulate this co-design as joint dynamic programming (joint-DP), a unified optimization over sensor geometry and a Bellman-optimal adaptive measurement policy. The outer hardware gradient is obtained through differentiable dynamic programming with a sharp Bellman maximum. A hierarchy of relaxations extends the framework from small discrete POMDPs to freeform photonic topologies with more than $10^5$ design pixels.
Figures
Forward citations
Cited by 2 Pith papers
-
Optically Incoherent Photonic Mutual Information
An end-to-end electromagnetic channel model shows point focusing is uniquely optimal for isotropic incoherent sources under a spectrum-flattening condition, while oversampled intensity detection favors interferometric...
-
End-to-end meta-imagers: Information-theoretic objectives and generalized focusing optima
For intensity-only detectors, optimal incoherent transfer matrices for Shannon and Fisher objectives are permutation matrices, so each source must focus onto a distinct detector.
Reference graph
Works this paper leans on
-
[1]
Satyajeet S Ahuja, Srinivasan Ramasubramanian, and Marwan Krunz. Srlg failure localization in optical networks.IEEE/ACM Transactions on Networking, 19(4):989–999, 2011
work page 2011
-
[2]
Anthony Atkinson, Alexander Donev, and Randall Tobias.Optimum experimental designs, with SAS, volume 34. OUP Oxford, 2007
work page 2007
-
[3]
Algorithms for network topology discovery using end-to-end measurements
Laurent Bobelin and Traian Muntean. Algorithms for network topology discovery using end-to-end measurements. In2008 International Symposium on Parallel and Distributed Computing, pages 267–
-
[4]
The convergence of a class of double-rank minimization algorithms 1
Charles George Broyden. The convergence of a class of double-rank minimization algorithms 1. general considerations.IMA Journal of Applied Mathematics, 6(1):76–90, 1970
work page 1970
-
[5]
Tian Bu, Nick Duffield, Francesco Lo Presti, and Don Towsley. Network tomography on general topolo- gies.ACM SIGMETRICS Performance Evaluation Review, 30(1):21–30, 2002
work page 2002
-
[6]
Simple network management protocol (SNMP)
Jeffrey D Case, Mark Fedor, Martin L Schoffstall, and James Davin. Simple network management protocol (SNMP). Technical report, 1989
work page 1989
-
[7]
Network Tomography: Recent Developments.Statistical Science, 19(3):499 – 517, 2004
Rui Castro, Mark Coates, Gang Liang, Robert Nowak, and Bin Yu. Network Tomography: Recent Developments.Statistical Science, 19(3):499 – 517, 2004
work page 2004
-
[8]
Network health and e-science in commercial clouds
Ryan Chard, Kris Bubendorfer, and Bryan Ng. Network health and e-science in commercial clouds. Future Generation Computer Systems, 56:595–604, 2016
work page 2016
-
[9]
Network tomography for internal delay estimation
Mark J Coates and Robert D Nowak. Network tomography for internal delay estimation. In2001 IEEE International Conference on Acoustics, Speech, and Signal Processing. Proceedings (Cat. No. 01CH37221), volume 6, pages 3409–3412. IEEE, 2001
work page 2001
-
[10]
Thomas H Cormen, Charles E Leiserson, Ronald L Rivest, and Clifford Stein.Introduction to algorithms. MIT press, 2022
work page 2022
-
[11]
Parameter orthogonality and approximate conditional inference
David Roxbee Cox and Nancy Reid. Parameter orthogonality and approximate conditional inference. Journal of the Royal Statistical Society: Series B (Methodological), 49(1):1–18, 1987
work page 1987
-
[12]
Quantum network tomography.IEEE Network, 38(5):114–122, 2024
Matheus Guedes De Andrade, Jake Navas, Saikat Guha, In` es Monta˜ no, Michael Raymer, Brian Smith, and Don Towsley. Quantum network tomography.IEEE Network, 38(5):114–122, 2024
work page 2024
-
[13]
Kiril Dichev, Fergal Reid, and Alexey Lastovetsky. Efficient and reliable network tomography in hetero- geneous networks using bittorrent broadcasts and clustering algorithms. InSC’12: Proceedings of the International Conference on High Performance Computing, Networking, Storage and Analysis, pages 1–11. IEEE, 2012
work page 2012
-
[14]
Network tomography from measured end-to-end delay covariance
Nick G Duffield and F Lo Presti. Network tomography from measured end-to-end delay covariance. IEEE/ACM Transactions On Networking, 12(6):978–992, 2004
work page 2004
-
[15]
A new approach to variable metric algorithms.The computer journal, 13(3):317–322, 1970
Roger Fletcher. A new approach to variable metric algorithms.The computer journal, 13(3):317–322, 1970
work page 1970
-
[16]
Resource allocation via graph neural networks in free space optical fronthaul networks
Zhan Gao, Mark Eisen, and Alejandro Ribeiro. Resource allocation via graph neural networks in free space optical fronthaul networks. InGLOBECOM 2020-2020 IEEE Global Communications Conference, pages 1–6. IEEE, 2020
work page 2020
-
[17]
Manuel Gessner, Augusto Smerzi, and Luca Pezz` e. Multiparameter squeezing for optimal quantum enhancements in sensor networks.Nature communications, 11(1):3817, 2020. 18
work page 2020
-
[18]
Netscope: Prac- tical network loss tomography
Denisa Ghita, Hung Nguyen, Maciej Kurant, Katerina Argyraki, and Patrick Thiran. Netscope: Prac- tical network loss tomography. In2010 Proceedings IEEE INFOCOM, pages 1–9. IEEE, 2010
work page 2010
-
[19]
Donald Goldfarb. A family of variable-metric methods derived by variational means.Mathematics of computation, 24(109):23–26, 1970
work page 1970
-
[20]
Saikat Guha, Tiju Cherian John, Zihao Gong, and Prithwish Basu. Quantum-enhanced quickest change detection of transmission loss.Physical Review Letters, 135(21):210801, 2025
work page 2025
-
[21]
Non- adaptive fault diagnosis for all-optical networks via combinatorial group testing on graphs
Nicholas JA Harvey, Mihai Patrascu, Yonggang Wen, Sergey Yekhanin, and Vincent WS Chan. Non- adaptive fault diagnosis for all-optical networks via combinatorial group testing on graphs. InIEEE INFOCOM 2007-26th IEEE International Conference on Computer Communications, pages 697–705. IEEE, 2007
work page 2007
-
[22]
Ting He, Chang Liu, Ananthram Swami, Don Towsley, Theodoros Salonidis, Andrei Iu Bejan, and Paul Yu. Fisher information-based experiment design for network tomography.ACM SIGMETRICS Performance Evaluation Review, 43(1):389–402, 2015
work page 2015
-
[23]
Cambridge University Press, 2021
Ting He, Liang Ma, Ananthram Swami, and Don Towsley.Network tomography: identifiability, mea- surement design, and network state inference. Cambridge University Press, 2021
work page 2021
-
[24]
Fredrik Korsb¨ ack, Lincoln Dale, and Dave McGaugh. Growing AWS internet peering with 400 GbE, 2023.������������������������������������������������������������� ������������������������������������������
work page 2023
-
[25]
Link budget analysis for free-space optical satellite networks
Jintao Liang, Aizaz U Chaudhry, Eylem Erdogan, and Halim Yanikomeroglu. Link budget analysis for free-space optical satellite networks. In2022 IEEE 23rd International Symposium on a World of Wireless, Mobile and Multimedia Networks (WoWMoM), pages 471–476. IEEE, 2022
work page 2022
-
[26]
Hybrid classical-quantum communication networks
Joseph M Lukens, Nicholas A Peters, and Bing Qi. Hybrid classical-quantum communication networks. Progress in Quantum Electronics, page 100586, 2025
work page 2025
-
[27]
Efficient identification of addi- tive link metrics via network tomography
Liang Ma, Ting He, Kin K Leung, Don Towsley, and Ananthram Swami. Efficient identification of addi- tive link metrics via network tomography. In2013 IEEE 33rd International Conference on Distributed Computing Systems, pages 581–590. IEEE, 2013
work page 2013
-
[28]
Stephen Nellis. Cisco and qunnect build quantum network using new york fiber optic cables.����������������������������������������������� ���������������������������������������������������������������������������������, February 2026. Accessed: 2026-04-17
work page 2026
-
[29]
Multiparameter gaussian quantum metrology.Physical Review A, 98(1):012114, 2018
Rosanna Nichols, Pietro Liuzzo-Scorpo, Paul A Knott, and Gerardo Adesso. Multiparameter gaussian quantum metrology.Physical Review A, 98(1):012114, 2018
work page 2018
-
[30]
Jorge Nocedal and Stephen J Wright.Numerical optimization. Springer, 2006
work page 2006
-
[31]
Leon Poutievski, Omid Mashayekhi, Joon Ong, Arjun Singh, Mukarram Tariq, Rui Wang, Jianan Zhang, Virginia Beauregard, Patrick Conner, Steve Gribble, et al. Jupiter evolving: transforming google’s datacenter network via optical circuit switches and software-defined networking. InProceedings of the ACM SIGCOMM 2022 Conference, pages 66–85, 2022
work page 2022
-
[32]
David F Shanno. Conditioning of quasi-newton methods for function minimization.Mathematics of computation, 24(111):647–656, 1970
work page 1970
-
[33]
Jack Sherman and Winifred J Morrison. Adjustment of an inverse matrix corresponding to a change in one element of a given matrix.The Annals of Mathematical Statistics, 21(1):124–127, 1950. 19
work page 1950
-
[34]
Passive network tomography using em algorithms
Yolanda Tsang, Mark Coates, and Robert Nowak. Passive network tomography using em algorithms. In2001 IEEE International Conference on Acoustics, Speech, and Signal Processing. Proceedings (Cat. No. 01CH37221), volume 3, pages 1469–1472. IEEE, 2001
work page 2001
-
[35]
Network delay tomography.IEEE Transactions on Signal Processing, 51(8):2125–2136, 2003
Yolanda Tsang, Mark Coates, and Robert D Nowak. Network delay tomography.IEEE Transactions on Signal Processing, 51(8):2125–2136, 2003
work page 2003
-
[36]
Network tomography: Estimating source-destination traffic intensities from link data
Yehuda Vardi. Network tomography: Estimating source-destination traffic intensities from link data. Journal of the American statistical association, 91(433):365–377, 1996
work page 1996
-
[37]
Lukas Velush. Boosting our connectivity with our own next-generation optical network, 2023.������������������������������������������� �����������������������������������������������������������������������
work page 2023
-
[38]
Xuchuang Wang, Matheus Guedes De Andrade, Guus Avis, Yu-Zhen Janice Chen, Mohammad Hajies- maili, and Don Towsley. Quantum network tomography for general topology with spam errors.arXiv preprint arXiv:2511.01074, 2025
-
[39]
Bowei Xi, George Michailidis, and Vijayan N Nair. Estimating network loss rates using active tomog- raphy.Journal of the American Statistical Association, 101(476):1430–1448, 2006
work page 2006
-
[40]
Yao Zhao, Yan Chen, and David Bindel. Towards unbiased end-to-end network diagnosis.IEEE/ACM Transactions on Networking, 17(6):1724–1737, 2009
work page 2009
-
[41]
A quantum speedup in localizing transmission loss change in optical networks
Yufei Zheng, Yu-Zhen Janice Chen, Prithwish Basu, and Don Towsley. A quantum speedup in localizing transmission loss change in optical networks. In2025 IEEE International Conference on Quantum Computing and Engineering (QCE), volume 1, pages 958–968. IEEE, 2025. 20 (a) det(� s�s)�det(� e�s) wrtη 1 andη 2. (b) Tr(� �1 e�s )�Tr(� �1 s�s ) wrtη 1 andη 2. Fig...
work page 2025
discussion (0)
Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.