REVIEW 5 cited by
The Hardware Lottery
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Hardware, systems and algorithms research communities have historically had different incentive structures and fluctuating motivation to engage with each other explicitly. This historical treatment is odd given that hardware and software have frequently determined which research ideas succeed (and fail). This essay introduces the term hardware lottery to describe when a research idea wins because it is suited to the available software and hardware and not because the idea is superior to alternative research directions. Examples from early computer science history illustrate how hardware lotteries can delay research progress by casting successful ideas as failures. These lessons are particularly salient given the advent of domain specialized hardware which make it increasingly costly to stray off of the beaten path of research ideas. This essay posits that the gains from progress in computing are likely to become even more uneven, with certain research directions moving into the fast-lane while progress on others is further obstructed.
Forward citations
Cited by 5 Pith papers
-
Spec Sheets Are Not Kernels: An ISA- and Source-Level Audit of INT8 Availability on NVIDIA Blackwell Ultra
On NVIDIA Blackwell Ultra, INT8 W8A8 is undeployable by default because the PTX ISA, CUTLASS, vLLM, and SGLang all lack a fifth-generation INT8 tensor-core path, despite the datasheet listing INT8 support.
-
Call for Action: towards the next generation of symbolic regression benchmark
An expanded SRBench comparison of 25 symbolic regression methods across 24 datasets shows no algorithm dominates all tasks and advocates for a living community benchmark.
-
Streaming DiLoCo with overlapping communication: Towards a Distributed Free Lunch
Streaming DiLoCo trains billion-parameter LLMs at data-parallel quality while cutting the inter-datacenter bandwidth by about two orders of magnitude via partial, overlapped, and 4-bit-quantized synchronization.
-
Scaling Deep Learning Training with MPMD Pipeline Parallelism
JaxPP introduces a user-defined MPMD pipeline schedule API on top of JAX/GSPMD and reports throughput gains up to 1.11x over SPMD training on H100 clusters.
-
Grounded world models in biological organisms and future embodied AI
Biological intelligence builds grounded world models through sensorimotor interaction and attaches language later, the opposite of language-first AI training.
Discussion (0). Continue with ORCID to comment.