Home › Census › get_similarity

get_similarity

Syntologyfunction-name censuscensus 2026-09-22battery b986f7e04d79all samples with this name

get_similarity: 3 implementations from 5 papers ran on one shared input (census 2026-09-22, battery b986f7e04d79); they produced 3 distinct outputs across 2 buckets, one shared input per bucket.

Identical values to six decimals (the recorded digest) on the shared input are agreement on those inputs, not equivalence. Implementations are compared only within one bucket, the positional (rank, kind, dtype) of each array argument; the argument name is not part of the key because the harness draws the shared array from (rank, kind) and casts it to the dtype, whatever the name; each bucket's shared input is fixed by that key, so members of one bucket saw bitwise-identical inputs under their own scalar arguments. A cluster is the set of members whose recorded output digest is identical. Nothing here says which computation a paper's method intended, and nothing reproduces a paper's results.

Not compared, and not in the tables or the counts above:

Bucket 1 of 2: arg 1: rank 2, kind float, dtype float64 · arg 2: rank 2, kind float, dtype float64

2 implementations from 3 papers share this bucket (rank, kind, dtype of each array argument, positional; each member's recorded signature, argument name included, is shown under it); 2 distinct outputs, largest cluster first. Values are the first 8 of the recorded output, flattened.

Cluster (same digest to six decimals)MembersShared output on this bucket's input
1 implementation
2 papers
af5570f5a181
one code sha held from 2 papers' repositories
all_baselines/fed-cluster/models/fedcluster.py b3d657e0
recorded X_c1:2/float/float64, X_c2:2/float/float64; c1='cluster_1', c2='cluster_2'
[0]
shape [] · float64
1 implementation
1 paper
92eb2826d9fa

extract_alignments.py 6e7c054c
recorded X:2/float/float64, Y:2/float/float64
[1, 0.440513, 0.869565, 0.380331, 0.440513, 1, 0.300492, 0.595765, …]
shape [4, 4] · float64 · ndarray

Bucket 2 of 2: arg 1: rank 4, kind float, dtype float32 · arg 2: rank 4, kind float, dtype float32 · arg 3: rank 3, kind int, dtype int64

1 implementation from 2 papers share this bucket (rank, kind, dtype of each array argument, positional; each member's recorded signature, argument name included, is shown under it); 1 distinct output, largest cluster first. Values are the first 8 of the recorded output, flattened.

Cluster (same digest to six decimals)MembersShared output on this bucket's input
1 implementation
2 papers
cf71868516e4
one code sha held from 2 papers' repositories
model/AMFormer.py 878c68be
recorded q:4/float/float32, s:4/float/float32, mask:3/int/int64
[0.999972, 1, 0.624481, 1, 1, 1, 0.334547, 0.228664, …]
shape [2, 1, 4, 4] · float32 · Tensor

Identical values to six decimals (the recorded digest) on the shared input are agreement on those inputs, not equivalence; where a cluster's members carry recorded values, the largest difference among them is shown under the cluster. Paper titles are the archive's archive 2025-07-28 where the paper is in the archive and the graph's where it was added by Syntology; papers with no page here are shown by their recorded paper id only. A paper count above the implementation count means one implementation (one code sha) is held from several papers' repositories and counts once. Per-sample status, licence and fingerprint records for each paper are on its paper page. JSON twin: /census/get-similarity.json.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections