Pith. sign in

REVIEW 1 cited by

HyperFields: Towards Zero-Shot Generation of NeRFs from Text

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2310.17075 v3 pith:M3RIG7E3 submitted 2023-10-26 cs.CV

classification cs.CV
keywords hyperfieldsnerfsscenesdynamictextcapabledistillationfinetuning
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

We introduce HyperFields, a method for generating text-conditioned Neural Radiance Fields (NeRFs) with a single forward pass and (optionally) some fine-tuning. Key to our approach are: (i) a dynamic hypernetwork, which learns a smooth mapping from text token embeddings to the space of NeRFs; (ii) NeRF distillation training, which distills scenes encoded in individual NeRFs into one dynamic hypernetwork. These techniques enable a single network to fit over a hundred unique scenes. We further demonstrate that HyperFields learns a more general map between text and NeRFs, and consequently is capable of predicting novel in-distribution and out-of-distribution scenes -- either zero-shot or with a few finetuning steps. Finetuning HyperFields benefits from accelerated convergence thanks to the learned general map, and is capable of synthesizing novel scenes 5 to 10 times faster than existing neural optimization-based methods. Our ablation experiments show that both the dynamic architecture and NeRF distillation are critical to the expressivity of HyperFields.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Any-to-3D Generation via Hybrid Diffusion Supervision

    cs.CV 2024-11 reject novelty 6.0 of 10

    XBind generates 3D objects from text, image, or audio prompts using ImageBind aligned embeddings and hybrid 2D/3D diffusion supervision.

Pith tools