{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/slotpi-physics-informed-object-centric","title":"SlotPi: Physics-informed Object-centric Reasoning Models","arxiv_id":"2506.10778","date":"2025-06-12","proceeding":null,"authors":["Jian Li","Wan Han","Ning Lin","Yu-Liang Zhan","Ruizhi Chengze","Haining Wang","Yi Zhang","Hongsheng Liu","Zidong Wang","Fan Yu","Hao Sun"],"abstract":"Understanding and reasoning about dynamics governed by physical laws through visual observation, akin to human capabilities in the real world, poses significant challenges. Currently, object-centric dynamic simulation methods, which emulate human behavior, have achieved notable progress but overlook two critical aspects: 1) the integration of physical knowledge into models. Humans gain physical insights by observing the world and apply this knowledge to accurately reason about various dynamic scenarios; 2) the validation of model adaptability across diverse scenarios. Real-world dynamics, especially those involving fluids and objects, demand models that not only capture object interactions but also simulate fluid flow characteristics. To address these gaps, we introduce SlotPi, a slot-based physics-informed object-centric reasoning model. SlotPi integrates a physical module based on Hamiltonian principles with a spatio-temporal prediction module for dynamic forecasting. Our experiments highlight the model's strengths in tasks such as prediction and Visual Question Answering (VQA) on benchmark and fluid datasets. Furthermore, we have created a real-world dataset encompassing object interactions, fluid dynamics, and fluid-object interactions, on which we validated our model's capabilities. The model's robust performance across all datasets underscores its strong adaptability, laying a foundation for developing more advanced world models.","url_abs":"https://arxiv.org/abs/2506.10778v1","url_pdf":"https://arxiv.org/pdf/2506.10778v1.pdf","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","row_kind":"abstracts"},"code_links":[{"paper_slug":"slotpi-physics-informed-object-centric","repo_url":"https://github.com/intell-sci-comput/slotpi","is_official":1,"mentioned_in_paper":1,"mentioned_in_github":0,"framework":"pytorch","reach":null}],"tasks":[{"task_slug":"object","task_name":"Object"},{"task_slug":"question-answering","task_name":"Question Answering"},{"task_slug":"visual-question-answering-1","task_name":"Visual Question Answering"},{"task_slug":"visual-question-answering","task_name":"Visual Question Answering (VQA)"}],"methods":[],"datasets_introduced":[],"methods_introduced":[],"results":[],"syntology":{"atlas_url":"https://app.syntology.ai/?focus=2506.10778","mcp":{"get_harvested_code_for_paper":{"arxiv_id":"2506.10778"}},"developers":"https://syntology.ai/developers","read_at":"2026-09-24T18:15:14+00:00","read_at_is":"when the build read Syntology's graph, not when any sample ran","claim":"Per-sample execution status on synthesized fixtures; not a correctness claim about the paper. Samples come from repositories linked to the paper, official or community; repo_kind says which.","repos":[{"provenance":"external:paperswithcode_snapshot_2025-07-28","url":"https://github.com/intell-sci-comput/slotpi","reach":null}],"summary":{"ran_honours":2,"ran_fixture":1},"by_repo_kind":{"official":{"samples":3,"ran":3,"repositories":1}},"repo_kind_vocabulary":{"official":"The archive marks this repository official for the paper","named_in_paper":"The archive records that the paper mentions this repository; it is not marked official","listed":"In the archive's code links for this paper, not marked official and not recorded as mentioned in the paper","found_in_text":"Syntology found this repository in the paper's own text; whether it is the authors' implementation is not asserted","community":"Not in the archive's code links for this paper; a community repository Syntology harvested"},"n_pointer_only_for_licence":3,"samples":[{"code_sha256_prefix":"1aadff67e8877271","entry":"build_pos_enc","repo":"intell-sci-comput/slotpi","repo_kind":"official","path":"models/diffusion_slot_CVDT.py","file_url":"https://github.com/intell-sci-comput/slotpi/blob/HEAD/models/diffusion_slot_CVDT.py","link_basis":"first_harvest_node","language":"python","status":"ran_honours","verification_level":1,"contract_check":"HONOURS","metamorphic_tier":"well_formed","behaviour_fingerprint":true,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"1aadff67e8877271"}},{"code_sha256_prefix":"1a9a0567874bbf3e","entry":"get_sin_pos_enc","repo":"intell-sci-comput/slotpi","repo_kind":"official","path":"models/diffusion_slot_CVDT.py","file_url":"https://github.com/intell-sci-comput/slotpi/blob/HEAD/models/diffusion_slot_CVDT.py","link_basis":"first_harvest_node","language":"python","status":"ran_honours","verification_level":1,"contract_check":"HONOURS","metamorphic_tier":"well_formed","behaviour_fingerprint":true,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"1a9a0567874bbf3e"}},{"code_sha256_prefix":"c081e2d896c9bd7b","entry":"modulate","repo":"intell-sci-comput/slotpi","repo_kind":"official","path":"models/diffusion_slot_CVDT.py","file_url":"https://github.com/intell-sci-comput/slotpi/blob/HEAD/models/diffusion_slot_CVDT.py","link_basis":"first_harvest_node","language":"python","status":"ran_fixture","verification_level":1,"contract_check":"RAISES","metamorphic_tier":"invariant","behaviour_fingerprint":false,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"c081e2d896c9bd7b"}}]},"arxiv_metadata":null,"syntology_extracted_results":null}