Pith. sign in

REVIEW 6 cited by

Dive into Deep Learning

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2106.11342 v5 pith:IVQHLKPH submitted 2021-06-21 cs.LG cs.AIcs.CLcs.CV

classification cs.LGcs.AIcs.CLcs.CV
keywords codelearningbookdeepinteractiveofferreaderstechnical
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

This open-source book represents our attempt to make deep learning approachable, teaching readers the concepts, the context, and the code. The entire book is drafted in Jupyter notebooks, seamlessly integrating exposition figures, math, and interactive examples with self-contained code. Our goal is to offer a resource that could (i) be freely available for everyone; (ii) offer sufficient technical depth to provide a starting point on the path to actually becoming an applied machine learning scientist; (iii) include runnable code, showing readers how to solve problems in practice; (iv) allow for rapid updates, both by us and also by the community at large; (v) be complemented by a forum for interactive discussion of technical details and to answer questions.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 6 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. TOBACO: Topology Optimization via Band-limited Coordinate Networks for Compositionally Graded Alloys

    cs.CE 2025-08 conditional novelty 7.0 of 10

    TOBACO maps a composition-gradation manufacturing limit to a neural-network bandwidth via Bernstein's inequality, making the constraint implicit in the design representation.

  2. RAMS: Residual-based adversarial-gradient moving sample method for scientific machine learning in solving partial differential equations

    cs.CE 2025-09 conditional novelty 6.0 of 10

    Treating training samples as trainable parameters and moving them along the residual's adversarial gradient improves accuracy across PINN and operator learning benchmarks.

  3. Token Statistics Transformer: Linear-Time Attention via Variational Rate Reduction

    cs.LG 2024-12 conditional novelty 6.0 of 10

    A variational reformulation of the MCR2 objective yields a linear-complexity attention operator, ToST, that matches transformer performance without computing pairwise token similarities.

  4. HashAttention: Semantic Sparsity for Faster Inference

    cs.LG 2024-12 conditional novelty 6.0 of 10

    HashAttention uses learned 32-bit bit-signatures to retrieve the pivotal context tokens for each query, enabling sparse attention with up to 16x token reduction and minimal average quality loss.

  5. Federated Testing (FedTest): A New Scheme to Enhance Convergence and Mitigate Adversarial Attacks in Federating Learning

    cs.LG 2025-01 conditional novelty 5.0 of 10

    FedTest lets users evaluate each other's models with local data and aggregates models by these scores, claiming faster convergence and better robustness to malicious users.

  6. Explaining in Diffusion: Explaining a Classifier Through Hierarchical Semantics with Text-to-Image Diffusion Models

    cs.CV 2024-12 conditional novelty 5.0 of 10

    DiffEx explains classifier decisions by using a vision-language model to build a hierarchical semantic corpus and a beam-search algorithm to rank which visual attributes, alone or in combination, most influence classi...

Pith tools