REVIEW 5 cited by
ToDMA: Large Model-Driven Massive Token Communications for Semantic Multiple Access
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Token communications (TokenCom) is an emerging generative semantic communication paradigm, where tokens serve as compact representation units across modalities. Their contextual dependencies can be exploited by pretrained large models for semantic recovery. In this paper, we propose token-domain multiple access (ToDMA), a large-model-driven semantic multiple access scheme for massive token communications. ToDMA integrates unsourced random access with context-aware token processing. It enables massive uncoordinated devices to transmit tokenized source representations over common uplink resources. Specifically, each token index is associated with a shared modulation codeword, exposing token-level structure to the receiver for context-aware recovery. At the receiver, compressed sensing is first employed to jointly detect active tokens and estimate their corresponding channel state information (CSI) from the superposed signals. The source token sequences are then reconstructed by exploiting the consistency of token-associated CSI across multiple token positions. In the presence of token collisions, some active tokens may remain unassigned, leading to missing entries in the reconstructed token sequences. To recover these tokens, candidate-restricted masked-token prediction is performed using pretrained contextual models, thereby leveraging token-level context to mitigate collision effects. Simulation results on both image and text transmission tasks demonstrate that ToDMA reduces access latency while maintaining favorable token recovery and semantic reconstruction quality, showing its scalability for semantic multiple access.
Forward citations
Cited by 5 Pith papers
-
Test-Time Scalable AI-RAN: Inference Time Allocation for Cell-Free MIMO
A new framework allocates inference time between test-time-scalable precoding and quantization AI modules in cell-free MIMO, with the optimal split depending on the temporal correlation of the channel.
-
Geometric Cross-Modal Token Selection for Latency-Constrained Multimodal Token Communication
Selecting tokens that lie inside multiple anchor-centric semantic grain regions improves multimodal VQA/AVQA accuracy under latency and erasure constraints compared with pairwise attention-based selection.
-
Wireless TokenCom: RL-Based Tokenizer Agreement for Multi-User Wireless Token Communications
Joint tokenizer/codebook selection, subchannel assignment, and beamforming for multi-user video TokenCom is posed as an MDP and solved by DQN for discrete choices and DDPG for beamforming, with simulated gains over H.265.
-
Text-Guided Token Communication for Wireless Image Transmission
A text-guided token transmission system using pre-trained image and text models outperforms a deep JSCC baseline on perceptual and semantic metrics, but relies on an assumption that text is available at the receiver.
-
Token Communication in the Era of Large Models: An Information Bottleneck-Based Approach
A unified token-based wireless communication framework combines an information-bottleneck-style tokenizer with a causal multimodal language model for joint understanding and generation.
Discussion (0). Continue with ORCID to comment.