Subset-Based Collective Decision-Making
- SubCDM is a framework where collective decisions are made from restricted subsets, enabling structured partial participation rather than full-scale aggregation.
- It spans several methodologies—including social choice proportionality, decentralized swarm consensus, partial-information macroscopes, and higher-order dynamic models—to meet different application needs.
- The approach offers practical benefits such as improved fairness, resource efficiency, and communication compression, while also highlighting domain-specific limitations to be addressed.
to=arxiv_search 天天中彩票双色球json {"query":"(Masařík et al., 2023) OR (Fuady et al., 1 Aug 2025) OR (Ramamoorthy et al., 2012) OR (Boes et al., 2021) OR (Alouf-Heffetz et al., 2022) OR (Zafeiris et al., 2016) OR (Brandl et al., 2021) OR (Fontan et al., 5 Sep 2025) OR (Goswami et al., 2014) OR (Gilbert et al., 2023)","max_results":10,"sort_by":"submittedDate"} code to=arxiv_search ՞նչjson {"query":"(Masařík et al., 2023)","max_results":3,"sort_by":"relevance"} code to=arxiv_search үрүмjson {"query":"A Generalised Theory of Proportionality in Collective Decision Making","max_results":5} code to=arxiv_search 】【。】【”】【json {"search_query":"A Generalised Theory of Proportionality in Collective Decision Making","max_results":5} code Subset-Based Collective Decision-Making (SubCDM) designates a class of collective-choice models in which either the outcome is a subset chosen under constraints or the decision process itself is restricted to subsets of agents, alternatives, issues, or observations. In the literature, the term does not denote a single formalism. Rather, it appears in several technically distinct settings: proportional representation over feasible subsets of candidates, decentralized swarm consensus using only a subset of robots, collective computation under subset-based information allotments, stochastic dynamics driven by sampled neighborhoods or pairwise subsets, and learning or optimization over subset-valued alternatives (Masařík et al., 2023, Fuady et al., 1 Aug 2025, Ramamoorthy et al., 2012).
1. Conceptual scope and recurring structure
Across these strands, SubCDM replaces unrestricted full-participation aggregation with structured partiality. The subset may be the chosen outcome, the active decision-making population, the locally visible data, or the interaction neighborhood used in an update rule.
| Strand | Subset object | Canonical formalism |
|---|---|---|
| Social choice | Feasible outcome subset | , BEJR/EJR/PJR |
| Swarm robotics | Decision-making robot subset | Hop-based or probabilistic recruitment |
| Communication complexity | Allotment subsets | Macroscope |
| Nonlinear and stochastic dynamics | Sampled pairs, neighborhoods, or hyperedges | Urn process, hypergeometric sampling, hypernetwork ODEs |
| Preference learning and optimization | Subsets as alternatives or feasible judgments | Robust ordinal regression; CDO as judgment aggregation |
A recurring misconception is that subset-based decision-making is simply a lossy approximation to full participation. The surveyed literature presents a more differentiated picture. In social choice, the subset is the outcome and proportionality is strengthened rather than weakened. In swarm robotics, restricting active participation is explicitly intended to preserve accuracy while reducing resource use. In communication-complexity models, subset structure and meta-information can reduce communication from input-scale disclosure to task-specific summaries. In nonlinear dynamics, higher-order subset interactions can change the bifurcation structure of the decision process itself.
2. Social-choice foundations: proportionality over cohesive voter subsets
The most systematic formalization of subset-based collective decision-making in social choice models a collective outcome as a feasible subset of items. Voters are , items are , and is a nonempty family of feasible sets, assumed closed under inclusion. Each voter has approval set , with utility 0. This framework unifies committee elections, public decisions, diversity-constrained selection, and collective scheduling (Masařík et al., 2023).
The central normative move is to define proportionality directly for arbitrary subsets 1, without predefining demographic groups. A subset 2 deserves 3 under Base Extended Justified Representation (BEJR) if, for every 4, either there exists 5 with 6 such that 7, or
8
An outcome 9 satisfies BEJR if every subset deserving 0 contains some voter 1 with 2. Extended Justified Representation (EJR) strengthens this by conditioning the claim on the actual outcome 3: for every 4, the same feasibility-or-proportional-overruling condition must hold. The framework also defines PJR and BPJR generalizations, with PJR using 5.
These axioms admit constructive rule adaptations. Proportional Approval Voting is generalized by maximizing
6
over the full feasibility family 7. Phragmén’s Sequential Rule is generalized through a continuous-load process in which voters earn budget at rate 8, candidates have price 9, and a candidate is purchased when its supporters collectively accumulate the price. Stable-priceability is generalized through candidate prices 0, unit-budget payments 1, support-only payments, exact coverage of selected items, no profitable deviation to an unselected item, and producer-stability.
The main structural result is exact. PAV satisfies EJR for all elections with matroid constraints, and for any non-matroid 2 there exists an election where PAV fails BEJR. Phragmén’s sequential rule satisfies PJR for elections with matroid constraints, and for any non-matroid 3 there exists an election where it fails BPJR. For stable-priceability, every stable-priceable outcome satisfies EJR under matroid constraints; if candidate prices are equal, then any stable-priceable outcome satisfies EJR under arbitrary constraints. This identifies matroid feasibility as the boundary at which these strong subset-based proportionality guarantees are available.
The framework recovers familiar settings as special cases. For 4, BEJR and EJR reduce to classic multiwinner proportionality. For public decisions with binary issues, 5 is partitioned into pairs and exactly one option per issue must be chosen. Diversity-constrained elections with disjoint attribute groups and per-group quotas form a matroid. Ranking and judgment-aggregation style constraints are generally non-matroid, and the necessity results explain why adapted PAV and Phragmén can fail there. The broader significance is that proportional fairness is no longer tied to a fixed-seat committee model; it becomes a property of arbitrary feasible subset selection.
3. Swarm-robotic SubCDM
In swarm robotics, SubCDM denotes a decentralized framework in which only a subset of robots performs the decision-making task. The studied problem is best-of-2 decision-making: determining the dominant environmental feature, black versus white tiles. The motivation is explicitly resource-oriented: fewer robots need to sense, move, and communicate; idle robots can be reallocated to other tasks; and performance stagnation or degradation in very large swarms can be avoided (Fuady et al., 1 Aug 2025).
The evaluated system uses 6 foot-bot robots in an 7 m arena with randomly distributed 8 cm black and white tiles. Communication is local range-and-bearing with 9 m, updates are asynchronous at 0 ticks/s, and task difficulty is varied by black-tile proportion from 1 to 2, corresponding to black:white ratios 3–4.
SubCDM operates in three phases: subset construction, collective decision-making using DMVD, and subset evaluation or adjustment. Role tenure is stabilized by sampling 5. Two local-information subset-construction strategies are studied. In the leader-based strategy, robots maintain shortest hop count to a leader via
6
and the decision-making subset is
7
In the distributed strategy, each robot maintains a local subset parameter 8 and joins with probability
9
Idle robots can relay up to three messages to maintain connectivity among randomly selected decision-makers.
Subset size is adaptive. In the leader-based variant, the leader waits until at least 0 opinions are collected; if the majority ratio exceeds 1 for 2 s, the decision at the current 3 is recorded, and the process continues until 4 consistent decisions are obtained. If no stable majority is achieved within 5 s, 6 is increased. In the distributed variant, each robot tracks confidence 7, initialized at 8, and updates it when hearing a neighboring decision-maker: 9 with 0. Every 1 decrement in 2 increments 3, thereby increasing 4.
The consensus protocol within 5 is DMVD. Exploration durations satisfy 6 with 7 s; a robot estimates local quality by 8, then disseminates for 9 with 0 s. Positive feedback arises because longer dissemination is associated with larger local quality estimates.
Simulation results show that both SubCDM variants maintain accuracy comparable to full-swarm DMVD while using fewer robots. The leader-based subset expands outward and plateaus around 1 in the 2-robot setup. Spatial organization differs sharply: Moran’s Index over 3 runs is 4 for leader-based and 5 for distributed selection. Convergence time increases with task difficulty for all methods; full-swarm DMVD is faster, whereas SubCDM trades speed for resource efficiency. Under reduced communication and faults, the distributed variant is more resilient, while the leader-based variant is sensitive to hierarchy disruption and leader faults. The paper explicitly does not provide formal proofs or bounds for convergence rates, error probabilities, or connectivity conditions; support is empirical via repeated ARGoS simulations and stability thresholds.
4. Partial-information and communication-complexity formulations
A different line of work interprets subset-based collective decision-making as collective computation under partial information. In the macroscope model, the global input is 6, the parties are 7, and each party 8 observes a subset 9. The allotment structure is 0, and a macroscope is the pair 1, where 2 is the global function to be computed (Ramamoorthy et al., 2012).
The model distinguishes two meta-information regimes. In the single-blind regime, each party knows the full allotment structure 3. In the double-blind regime, each party knows only its own subset indices and values. Communication is one-round simultaneous broadcast on a blackboard, and the cost is total bits transmitted. This isolates how subset overlap and knowledge of overlap affect the complexity of collective decisions.
General bounds are sharp. Every single-blind macroscope on 4 bits has a protocol of cost 5, and this is optimal. Every 6-player double-blind macroscope on 7 bits has a protocol with cost at most 8. The single-blind upper bound is achieved by responsibility assignment: for each index 9, the lowest-index party holding 0 broadcasts 1. The double-blind upper bound requires each party to reveal both its held indices and the values it sees.
For specific tasks, subset structure can lower communication dramatically. For 2-ary constancy detection, if 3 is the intersection graph over parties and 4 is its number of connected components, then
5
optimal up to factor 6. In the double-blind regime,
7
For Boolean step-function detection,
8
For approximate averaging with error tolerance 9, single-blind knowledge of duplication counts 00 yields
01
while there exist 02-player double-blind instances with
03
The central lesson is that subset overlap helps only when it is known. Connected overlaps can compress communication to component summaries, and knowledge of duplication factors prevents double counting. Without such meta-information, the same subset structure becomes a source of uncertainty that must be communicated away.
5. Stochastic, higher-order, and neighborhood-based dynamics
Several models treat subset-based collective decisions as emergent stochastic or nonlinear dynamics rather than explicit optimization. In an urn-based process for social choice, an urn contains balls labeled by alternatives, a random voter compares two sampled labels, the preferred label replaces the losing one, and with probability 04 a random mutation relabels a ball uniformly. The urn state 05 is a Markov chain on the discrete simplex, and the expected vector field is
06
where 07 is the skew-symmetric majority-margin matrix. The main theorem states that for sufficiently small 08 and large enough 09, the urn state spends at least a 10 fraction of time within 11 of some maximal lottery 12, and the probability of being within 13 increases exponentially in 14 (Brandl et al., 2021).
A different higher-order formulation models opinion dynamics on hypernetworks. Agents have states 15, pairwise interactions are encoded by 16, and 17-way interactions by 18. For 19,
20
Here 21 is a social-effort bifurcation parameter. Without higher-order terms, the system undergoes a symmetric pitchfork at
22
With 23-way interactions, Lyapunov–Schmidt reduction yields
24
so the pitchfork is unfolded into a saddle-node plus pitchfork, producing a bistable interval 25. In that interval, the community may remain in deadlock or converge to a nontrivial decision depending on initial conditions (Fontan et al., 5 Sep 2025).
A third model studies well-mixed binary-opinion swarms in which an agent updates after consulting a neighborhood subset of odd size 26, sampled without replacement. If 27 of 28 agents hold opinion 29, then the number 30 of 31 agents in the sampled subset is hypergeometric: 32 With majority/minority rule weights 33 and noise level 34, the macroscopic drift in 35 is
36
For majority-dominated rule sets, the dynamics exhibit stable fixed points near 37 without noise and interior stable points when noise is added; for all-minority rules, 38 becomes the single stable fixed point (Goswami et al., 2014).
Taken together, these models show that subset interactions are not merely a sparsified version of pairwise averaging. They can induce maximal-lottery approximation, exact finite-population drift laws, or bifurcation changes that are absent in pairwise-only systems. This suggests that the combinatorics of which subsets interact is itself a control parameter of collective decision formation.
6. Preference learning, optimization, and institutional design
SubCDM also appears in methods that learn preferences over subsets, optimize feasible subset outcomes, or deliberately choose a deciding subset. In robust ordinal regression for subset comparisons, a subset 39 is represented by its indicator vector and evaluated by an interaction-aware utility
40
where 41 is a learned support of interacting coalitions. Preference data induce a polyhedron 42, and robust dominance requires unanimity across all simplest supports and all consistent parameter vectors. Degree minimization is polynomial via the kernel
43
whereas cardinality, weighted-size, and lexicographic support minimization are NP-hard and addressed by mixed-integer programming. On the IMDb data reported in the paper, ORD attains Prediction Rate 44, Precision 45, Recall 46, and 47, versus 48 values of 49 for LR, 50 for SVM, and 51 for KNN (Gilbert et al., 2023).
A complementary optimization framework represents collective discrete optimisation as judgment aggregation with weighted issues. An agenda 52, integrity constraints 53, agent ballots 54, and a set scoring function 55 define modular rules
56
57
and ranked variants. This framework subsumes approval-based participatory budgeting, collective spanning trees, collective scheduling, and multiwinner rules. The paper proves, among other equivalences, that the median rule is equivalent to 58, the weighted median rule is equivalent to 59, and 60 is equivalent to Ranked Agenda. It also gives an ILP implementation, including a single-commodity-flow formulation for spanning trees (Boes et al., 2021).
Institutional SubCDM appears explicitly in the problem of appointing a subset of agents to vote on behalf of the whole group under uncertainty. In the Denying Access Problem (DAP), one removes at most 61 agents so that the remaining subset’s majority decisions are guaranteed to coincide with the objectively correct majority outcomes of the full group on as many issues as possible. DAP is NP-complete even in the one-dimensional proposal space, yet fixed-parameter tractable in both 62 and 63; the 64-parameterized result uses ILP over agent types in 65. In a radical one-dimensional domain, the paper gives polynomial-time algorithms for subset appointment, education, and constrained delegation, showing that endogenous guru sets under delegation can function as effective deciding subsets (Alouf-Heffetz et al., 2022).
Group design for multidimensional decisions adds a further institutional layer. A complex problem is decomposed into 66 independent sub-problems, group members have competence matrix 67, proposal quality is 68, and evaluations satisfy
69
Fitness is 70, with competence cost 71. The main qualitative result is that the best performing groups have at least one specialist for each sub-problem, but specialists also need some insight into the other sub-problems. Empirical analysis on ISI Web of Science data reports that, across nine trends, median citations increase with authors’ average interdisciplinarity, measured by the Shannon entropy of subject classes in their references (Zafeiris et al., 2016).
These approaches are methodologically diverse, but they share a design principle: subset structure can be learned, optimized, or institutionally imposed rather than passively inherited.
7. Limitations and open directions
The main limitations are domain-specific and technically consequential. In proportional social choice, the positive characterizations for adapted PAV and Phragmén stop exactly at matroid feasibility; under non-matroid constraints, BEJR or PJR can fail, and weighted-candidate settings such as participatory budgeting remain problematic for exact EJR/PJR guarantees (Masařík et al., 2023). In swarm robotics, convergence and robustness claims are empirical rather than theorem-level, and the leader-based variant is sensitive to reduced communication and leader faults, while the distributed variant depends on relay forwarding with a cap of three messages per relay (Fuady et al., 1 Aug 2025).
For partial-information macroscopes, the one-round deterministic model leaves open the effect of multiple rounds, graph-constrained communication, randomized protocols, and intermediate meta-information regimes between single-blind and double-blind (Ramamoorthy et al., 2012). In higher-order dynamics, heterogeneous nonlinearities, antagonistic interactions, noise, asynchronous updates, time-varying hypernetworks, and hyperedges of order greater than three are identified as extensions that may alter stability and bifurcation structure (Fontan et al., 5 Sep 2025). In robust ordinal regression, lexicographic sparsity selection is NP-hard and robust unanimity is intentionally conservative, often yielding abstention when the data do not support reliable prediction (Gilbert et al., 2023). In collective discrete optimisation as judgment aggregation, modular ILP formulations are expressive but can be outperformed by specialized algorithms, and strategyproofness and general proportionality notions are not resolved (Boes et al., 2021).
A final conceptual point follows from the literature as a whole. SubCDM is best understood not as a single algorithmic family but as a structural idea: collective decisions can be mediated by subsets, and the mathematical consequences depend on what those subsets represent—feasible outcomes, cohesive coalitions, active robots, sampled alternatives, observed indices, or specialized task domains. The technical results surveyed here show that subset structure can yield stronger fairness axioms, lower communication, lower resource use, richer nonlinear dynamics, or more reliable preference inference, but only under correspondingly specific assumptions about feasibility, information, or interaction topology.