← back to paper
arxiv: 2607.04819 · 2 revisions
Layer-Parallel Inference Reduces Encrypted Nonlinear Depth in Transformers