{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/arxiv-2602-22555","title":"Autoregressive Visual Decoding from EEG Signals","arxiv_id":"2602.22555","date":"2026-02-26","proceeding":null,"authors":["Sicheng Dai","Hongwang Xiao","Shan Yu","Qiwei Ye"],"abstract":"Electroencephalogram (EEG) signals have become a popular medium for decoding visual information due to their cost-effectiveness and high temporal resolution. However, current approaches face significant challenges in bridging the modality gap between EEG and image data. These methods typically rely on complex adaptation processes involving multiple stages, making it hard to maintain consistency and manage compounding errors. Furthermore, the computational overhead imposed by large-scale diffusion models limit their practicality in real-world brain-computer interface (BCI) applications. In this work, we present AVDE, a lightweight and efficient framework for visual decoding from EEG signals. First, we leverage LaBraM, a pre-trained EEG model, and fine-tune it via contrastive learning to align EEG and image representations. Second, we adopt an autoregressive generative framework based on a \"next-scale prediction\" strategy: images are encoded into multi-scale token maps using a pre-trained VQ-VAE, and a transformer is trained to autoregressively predict finer-scale tokens starting from EEG embeddings as the coarsest representation. This design enables coherent generation while preserving a direct connection between the input EEG signals and the reconstructed images. Experiments on two datasets show that AVDE outperforms previous state-of-the-art methods in both image retrieval and reconstruction tasks, while using only 10% of the parameters. In addition, visualization of intermediate outputs shows that the generative process of AVDE reflects the hierarchical nature of human visual perception. These results highlight the potential of autoregressive models as efficient and interpretable tools for practical BCI applications.","url_abs":"https://arxiv.org/abs/2602.22555","url_pdf":"https://arxiv.org/pdf/2602.22555","source":{"archive":null,"snapshot":"2025-07-28","note":"not in the Papers with Code archive (frozen at the snapshot)","row_kind":"graph","title_abstract_authors_date":"arXiv metadata, CC0 1.0 (https://info.arxiv.org/help/license)"},"code_links":[],"tasks":[],"methods":[],"datasets_introduced":[],"methods_introduced":[],"results":[],"syntology":{"atlas_url":"https://app.syntology.ai/?focus=2602.22555","mcp":{"get_harvested_code_for_paper":{"arxiv_id":"2602.22555"}},"developers":"https://syntology.ai/developers","read_at":"2026-09-24T18:15:14+00:00","read_at_is":"when the build read Syntology's graph, not when any sample ran","claim":"Per-sample execution status on synthesized fixtures; not a correctness claim about the paper. Samples come from repositories linked to the paper, official or community; repo_kind says which.","repos":[{"mentioned_in_github":null,"is_official":null,"provenance":"deterministic:regex_extraction","mentioned_in_paper":null,"url":"https://github.com/ddicee/avde","reach":null}],"summary":{"ran":4,"unverified":3},"by_repo_kind":{"found_in_text":{"samples":7,"ran":4,"repositories":1}},"repo_kind_vocabulary":{"official":"The archive marks this repository official for the paper","named_in_paper":"The archive records that the paper mentions this repository; it is not marked official","listed":"In the archive's code links for this paper, not marked official and not recorded as mentioned in the paper","found_in_text":"Syntology found this repository in the paper's own text; whether it is the authors' implementation is not asserted","community":"Not in the archive's code links for this paper; a community repository Syntology harvested"},"n_pointer_only_for_licence":7,"samples":[{"code_sha256_prefix":"3b25c0ed7a996276","entry":"ClipLoss","repo":"ddicee/avde","repo_kind":"found_in_text","path":"models/labram.py","file_url":"https://github.com/ddicee/avde/blob/HEAD/models/labram.py","link_basis":"first_harvest_node","language":"python","status":"ran","verification_level":1,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"3b25c0ed7a996276"}},{"code_sha256_prefix":"84590d675e94d536","entry":"DropPath","repo":"ddicee/avde","repo_kind":"found_in_text","path":"models/labram.py","file_url":"https://github.com/ddicee/avde/blob/HEAD/models/labram.py","link_basis":"first_harvest_node","language":"python","status":"ran","verification_level":1,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":true,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"84590d675e94d536"}},{"code_sha256_prefix":"7e3b5cc2526c4ffd","entry":"PatchEmbed","repo":"ddicee/avde","repo_kind":"found_in_text","path":"models/labram.py","file_url":"https://github.com/ddicee/avde/blob/HEAD/models/labram.py","link_basis":"first_harvest_node","language":"python","status":"ran","verification_level":1,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"7e3b5cc2526c4ffd"}},{"code_sha256_prefix":"6b2453aa07d3fee2","entry":"TemporalConv","repo":"ddicee/avde","repo_kind":"found_in_text","path":"models/labram.py","file_url":"https://github.com/ddicee/avde/blob/HEAD/models/labram.py","link_basis":"first_harvest_node","language":"python","status":"ran","verification_level":1,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":true,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"6b2453aa07d3fee2"}},{"code_sha256_prefix":"b89948f53a64da1d","entry":"Block","repo":"ddicee/avde","repo_kind":"found_in_text","path":"models/labram.py","file_url":"https://github.com/ddicee/avde/blob/HEAD/models/labram.py","link_basis":"first_harvest_node","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"b89948f53a64da1d"}},{"code_sha256_prefix":"5387bdee61e77dca","entry":"NeuralTransformer","repo":"ddicee/avde","repo_kind":"found_in_text","path":"models/labram.py","file_url":"https://github.com/ddicee/avde/blob/HEAD/models/labram.py","link_basis":"first_harvest_node","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"5387bdee61e77dca"}},{"code_sha256_prefix":"689bf1cf525e2160","entry":"gather_features","repo":"ddicee/avde","repo_kind":"found_in_text","path":"models/labram.py","file_url":"https://github.com/ddicee/avde/blob/HEAD/models/labram.py","link_basis":"first_harvest_node","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"689bf1cf525e2160"}}]},"arxiv_metadata":{"licence":"arXiv metadata, CC0 1.0 (https://info.arxiv.org/help/license)","fields":["title","abstract","authors","date"],"primary_category":"cs.LG","source":"arxiv_2026.jsonl"},"syntology_extracted_results":null}