Browse State-of-the-Art › Collaborative Inference
Collaborative Inference
19 papers with code · 0 benchmarks · 0 datasets archive 2025-07-28
In collaborative inference, a single inference task is performed by multiple models distributed on two or more (typically resource-constrained IoT) devices
Description from the archive archive 2025-07-28.
Benchmarks archive 2025-07-28
No benchmark for this task in the archive.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
No dataset record in the archive lists this task.
Subtasks archive 2025-07-28
No subtask under this task in the archive's task tree.
Parent tasks archive 2025-07-28
Most implemented papers archive 2025-07-28
19 shown of 19 papers with code (68 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
30 Jul 2024 4 repositories listedWe consider collaborative inference at the wireless edge, where each client's model is trained independently on its local dataset.
-
28 Sep 2022 3 repositories listedThe cost efficiency of model inference is critical to real-world machine learning (ML) applications, especially for delay-sensitive tasks and resource-limited devices.
-
2 Sep 2022 2 repositories listedHowever, these techniques have innate limitations: offloading is too slow for interactive inference, while APIs are not flexible enough for research that requires access to weights, attention or logits.
-
8 Oct 2021 2 repositories listed Syntology ran 8 of 14 samples · 6 unverifiedSelf-supervised learning has been shown to be very effective in learning useful representations, and yet much of the success is achieved in data types such as images, audio, and text.
-
2 Nov 2016 2 repositories listedWe propose Dual Attention Networks (DANs) which jointly leverage visual and textual attention mechanisms to capture fine-grained interplay between vision and language.
-
1 Mar 2025 1 repository listed Syntology ran 1 of 1 samples · 0 unverifiedBy locally encoding raw data into intermediate features, collaborative inference enables end users to leverage powerful deep learning models without exposure of sensitive raw data to cloud servers.
-
7 Feb 2025 1 repository listed Syntology ran 0 of 1 samples · 1 unverifiedLearning informative representations of phylogenetic tree structures is essential for analyzing evolutionary relationships.
-
4 Feb 2025 1 repository listed Syntology ran 1 of 2 samples · 1 unverified · 2 pointer-only (licence)This allows the router to learn to predict token-level routing scores and make routing decisions based on both the current token and the future impact of its decisions.
-
30 Dec 2024 1 repository listedFurther, we demonstrate through experiments, leveraging open-source LLMs for the implementation of distributed MoA, that certain MoA configurations produce higher-quality responses compared to others, as evaluated on…
-
GREAT: Geometry-Intention Collaborative Inference for Open-Vocabulary 3D Object Affordance Grounding29 Nov 2024 1 repository listedIn light of this, we propose GREAT (GeometRy-intEntion collAboraTive inference) for Open-Vocabulary 3D Object Affordance Grounding, a novel framework that mines the object invariant geometry attributes and performs…
-
12 Jul 2024 1 repository listed Syntology ran 9 of 10 samples · 1 unverifiedIn this paper, we propose a Global-Local Collaborative Scheme (GLIS) for the lidar-based OVD task, which contains a local branch to generate object-level detection result and a global branch to obtain scene-level global…
-
20 Jun 2024 1 repository listed Syntology ran 1 of 1 samples · 0 unverifiedBy employing an large LLM for inference and a small LLM for output, we achieve an average 37% reduction in response latency, alongside a 4.
-
23 Apr 2024 1 repository listed Syntology ran 7 of 8 samples · 1 unverifiedSpecifically, we propose a cooperative framework, Generalist-Specialist Collaboration (GSCo), which consists of two stages, namely the construction of GFM and specialists, and collaborative inference on downstream tasks.
-
23 Feb 2024 1 repository listedTherefore, instead of employing the partitioning strategy, our framework utilizes a lightweight ViT model on the edge device, with the server deploying a complicated ViT model.
-
9 May 2023 1 repository listedWe discuss the necessity, challenges, and solution approaches for extending existing work on classical edge computing to integrate QPUs.
-
15 Jul 2022 1 repository listedTo compensate, the method introduces a plug-in interactive block to allow attention transfer from the client-side by producing a feature mask.
-
7 Jun 2022 1 repository listedThe success of deep neural networks (DNNs) is heavily dependent on computational resources.
-
24 May 2022 1 repository listedIn this paper, we study the multi-agent collaborative inference scenario, where a single edge server coordinates the inference of multiple UEs.
-
19 Feb 2021 1 repository listedThis paper presents PRICURE, a system that combines complementary strengths of secure multi-party computation (SMPC) and differential privacy (DP) to enable privacy-preserving collaborative prediction among multiple…
Syntology lines on 7 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections