Diplomacy & Coalitions · Review

Try to break every arrow.

No category is scored from the other side's deficit alone. Partner action is credited only when observed and attributed to the system or partner that supplied it. Current U.S. policy shocks are treated as correlated rather than independent observations. AIIB, World Bank, BRICS, NATO, and G7 output is not assigned wholesale to Beijing or Washington. Missing PRC data lowers confidence and never counts as nonperformance.

Internal adversarial review

This is a structured second-pass challenge and sensitivity review, not external peer review or government validation.

Cross-cutting alternatives

Six explanations compete with system advantage.

They are tested across categories so coalition behavior is not automatically credited to Washington or Beijing.

DC-A01

Threat and interest alignment

Coalitions form and act because partners independently perceive the same threat or benefit, not because either leading power has superior diplomatic capability.

DC-A02

Material-power substitution

Military protection, market access, finance, or coercion produces participation and outcomes; diplomacy mainly communicates material incentives generated in other domains.

DC-A03

Partner hedging

States join multiple institutions to preserve options, status, or access, so membership and favorable votes are not durable alignment with either system.

DC-A04

Institutional function and age

Differences in delivery reflect mandate, legal authority, maturity, and issue type rather than U.S. or PRC leadership.

DC-A05

Domestic-politics constraint

Elections, legislatures, bureaucratic capacity, public opinion, and elite turnover explain commitment conversion and durability more than diplomatic-system quality.

DC-A06

Visibility and selection bias

Open U.S. systems disclose failures while opaque PRC systems disclose successes; public research also overselects famous coalitions and observable formal activity.

Category rulings

What would change each judgment.

Every current call retains contrary evidence, reasons against stronger alternatives, and a concrete sensitivity rule.

01 · Diplomatic reach & expertiseContested

RulingThe global position is Contested. The formal networks are effectively equal in size, with the PRC stronger in Africa, East Asia, and the Pacific and the United States stronger in Europe, North and Central America, and South Asia. Current Asian evidence favors the PRC, while both systems show substantial professionalization machinery. No common public record measures mission staffing, language match, tour continuity, access quality, warning, and crisis-channel performance, so the category cannot support an edge at more than low comparative confidence.

Strongest counterevidence

  • A three-post global difference is not operationally meaningful without staff and function weights.
  • The PRC's much stronger East Asian footprint and 2025 regional diplomatic influence prevent a U.S. global edge.
  • Historical U.S. language and staffing gaps show that a mature training system does not guarantee qualified occupancy.
  • PRC ambassador professionalization does not reveal mission-wide expertise, continuity, or access quality.

Why not another call

  • Why not us partner edge — The PRC has the slightly larger global network, stronger coverage in several priority regions, record Asian diplomatic-influence evidence, and a professionalized ambassador corps.
  • Why not prc centered edge — The United States has near-equal global coverage, stronger reach in Europe and the Americas, a deeply institutionalized language-training system, and dense partner institutions; the PRC's regional gains do not generalize globally.
  • Why not either lead — Neither side demonstrates a global advantage across coverage, expertise, access, continuity, warning, and crisis readiness under common measures.

Rival explanations applied

DC-A01 Material. Regional security and economic priorities explain post placement and diplomatic tempo.

DC-A02 Material. Market size, aid, and military presence can generate access credited superficially to diplomacy.

DC-A06 Decisive for confidence. U.S. staffing failures are disclosed through oversight while PRC personnel and failure data are sparse.

Sensitivity and revision conditions

to us partner edge

Requires current matched evidence that U.S. posts have higher qualified-fill, continuity, access, warning, and crisis-channel performance across multiple regions.

to prc centered edge

Requires the PRC's Asian influence and regional post advantages to convert into superior recurring access and problem-solving across a representative global sample.

to insufficient comparative basis

Revert to unscored if the category is required to be decided only from full mission-level expertise and outcome data rather than a profile combining reach and bounded performance evidence.

current call robust to
  • Removing meeting counts and ambassador rank from decisive evidence
  • Treating the 274-to-271 post difference as substantively equal
  • Separating regional from global findings
current call not robust to
  • A new matched mission staffing and access dataset
  • Large undisclosed 2024-2026 post closures or workforce changes
  • A rule that forbids component-profile adjudication when expertise remains unscored

Fact confidence: High · Comparative confidence: Low · Observability: limited

Return to the position →
02 · Informal reach & accessContested

RulingThe category is Contested at low confidence. The United States operates a large, plural, long-lived ecosystem of exchanges, alumni, educational, civic, professional, business, and subnational ties with documented participant and collaboration effects. The PRC operates an unusually taskable party-state network that reaches governing parties, opposition actors, and power brokers beyond foreign ministries, with observed though generally moderate effects in Southeast Asia. The networks differ too much in sponsor, target, disclosure, autonomy, and outcome measures to justify an edge; each has a distinct advantage that the other does not replicate.

Strongest counterevidence

  • More than half of the parties in the 2002-2017 ID-CPC dataset had only one or two observed contacts.
  • U.S. alumni prominence is heavily affected by participant selection and does not prove favorable policy behavior.
  • Southeast Asian research characterizes CPC party diplomacy as a moderate status-quo multiplier, not a structural transformation.
  • Central taskability can impose legitimacy and partner-autonomy costs, while plural U.S. ties are less controllable and therefore cannot always be mobilized.

Why not another call

  • Why not us partner edge — The CPC's party network supplies centrally directed access to influential ruling and opposition actors at a scale the U.S. government does not reproduce.
  • Why not prc centered edge — The U.S. plural ecosystem is older, broader across social sectors, more resilient to centralized policy failure, and supported by evaluations showing continuing collaboration and leadership networks.
  • Why not either lead — Neither system has a matched record of relationship-to-policy conversion, and each network's apparent strength is partly a consequence of unlike institutional form.

Rival explanations applied

DC-A02 Strong. Exchange, training, business, and party relationships often track material opportunities rather than autonomous diplomatic influence.

DC-A03 Strong. Participants may join both systems' networks to hedge, gain skills, or preserve access without policy alignment.

DC-A05 Material. Regime type and party dominance shape whether party ties or plural civic networks produce access.

DC-A06 Decisive for confidence. Private ties and failures are underreported; U.S. evaluations and CPC activity records use incompatible outcomes.

Sensitivity and revision conditions

to us partner edge

Requires matched cohorts showing greater U.S. network persistence, reciprocal access, warning, and policy conversion after leadership change.

to prc centered edge

Requires current entity-level evidence that CPC relationships convert into agenda entry or policy more often without prohibitive autonomy and legitimacy costs.

to insufficient comparative basis

Revert to unscored if unlike sponsor and outcome forms are deemed too heterogeneous for a profile judgment.

current call robust to
  • Excluding audience and contact totals from outcome credit
  • Treating U.S. and PRC networks as plural and party-state systems rather than unitary blocs
  • Retaining regional asymmetry
current call not robust to
  • A current matched network panel with verified policy outcomes
  • Large 2025-2026 program closures or revived exchange portfolios
  • Systematic partner-side evidence of coercion or nonreciprocity

Fact confidence: Moderate · Comparative confidence: Low · Observability: limited

Return to the position →
03 · Coalition formingContested

RulingCoalition forming is Contested. The U.S./partner system has the stronger observed portfolio of function-specific coalitions that include pivotal high-capability states and assign costly, differentiated roles. The PRC-centered system has the stronger recent record of expanding Global South forums and attracting members and partners seeking greater voice, while the PRC position also aligns with much broader coalitions on many UN General Assembly votes. Neither strength dominates the full category: threat and material incentives explain much U.S. participation, while BRICS and vote breadth do not establish PRC command, role coverage, activation, or delivery.

Strongest counterevidence

  • U.S.-anchored security coalitions benefit from direct threat and material military dependence.
  • BRICS members constrained expansion choices and pursue distinct visions, so expansion cannot be assigned to Beijing alone.
  • UN vote alignment is not sponsorship, persuasion, or meaningful role assignment.
  • Large coalitions can be shallow, while narrow coalitions can fill every decisive role.

Why not another call

  • Why not us partner edge — BRICS expansion, PRC-centered forum demand, party networks, and recent UN vote-coalition breadth are genuine positive coalition-forming evidence.
  • Why not prc centered edge — The U.S./partner system repeatedly assembles pivotal actors into specific high-cost roles rather than only broad membership or aligned declarations.
  • Why not either lead — Partner agency, hedging, threat alignment, and role-versus-breadth tradeoffs prevent either system from dominating across security, development, rules, and regions.

Rival explanations applied

DC-A01 Strong. Shared threats explain much U.S.-anchored security coalition formation.

DC-A02 Strong. Protection, market access, finance, and sanctions risk shape participation in both systems.

DC-A03 Strong. BRICS and other memberships often represent hedging and voice-seeking rather than alignment.

DC-A06 Material. Public records observe formal coalition entry more readily than private failed recruitment or side payments.

Sensitivity and revision conditions

to us partner edge

Requires a predeclared cross-issue sample showing superior pivotal-role coverage and voluntary participation after controlling for threat and material dependence.

to prc centered edge

Requires BRICS, FOCAC, and UN breadth to convert into verified partner-defined roles across multiple consequential issues without coercion or one-sided dependence.

current call robust to
  • Removing raw membership and GDP totals
  • Separating formation from activation and delivery
  • Treating members as autonomous actors
current call not robust to
  • A representative failed-recruitment denominator
  • Common role maps across security and non-security coalitions
  • Evidence that current U.S. policy discontinuity causes pivotal partner defections

Fact confidence: High · Comparative confidence: Moderate · Observability: partial

Return to the position →
04 · Commitment credibilityU.S. / partner edge

RulingThe U.S./partner system has a low-confidence edge. It sustains a broader observed architecture of ratified, geographically explicit security obligations connected to domestic authorization, standing institutions, plans, forces, exercises, and budgets. PRC credibility is real in development-finance delivery, FOCAC follow-up, and the durable China-DPRK treaty, but its security obligations are fewer and many political commitments are less specific or less observable. The edge is narrow: U.S. withdrawal from the initially implemented JCPOA, current foreign-aid disruption, international-organization withdrawals, arrears, and WTO obstruction materially damage continuity; the PRC's Hong Kong record and opaque commitment universe similarly preclude a PRC advantage.

Strongest counterevidence

  • Updated alliance research finds meaningful historical violation and coding sensitivity rather than automatic treaty reliability.
  • The U.S. withdrew from the JCPOA after initial IAEA-verified implementation.
  • The 2025 foreign-aid pause and 2025-2026 international withdrawals create a current cross-program continuity shock.
  • PRC development finance includes large volumes of completed and repaired projects and partner demand, refuting an empty-promise thesis.
  • The United Kingdom's continuing Joint Declaration noncompliance assessment is serious PRC counterevidence even though Beijing disputes it.

Why not another call

  • Why not us partner lead — Current U.S. reversals, arrears, aid interruption, and alliance-reliability uncertainty are too consequential; formal specificity cannot substitute for fulfillment.
  • Why not contested — The U.S./partner system has the broader positive record of explicit obligations already connected to authorization, resources, exercises, and live activation, not merely a PRC evidence deficit.
  • Why not prc centered edge or lead — PRC development delivery is strong counterevidence, but no comparable multi-region portfolio matches the specificity, standing resources, and observed activation of U.S.-anchored commitments; major PRC compliance disputes remain.

Rival explanations applied

DC-A01 Strong. Treaties are selected where interests already align, and changing interests explain much fulfillment.

DC-A02 Strong. Military and financial resources make promises credible; this category credits only their conversion into specific, authorized commitment.

DC-A05 Strong. Elections and domestic law drive U.S. discontinuity, while PRC centralization may increase short-run continuity but reduce transparent accountability.

DC-A06 Decisive for confidence. U.S. obligations and failures are more observable; missing PRC evidence is never coded as nonperformance.

Sensitivity and revision conditions

down to contested

Move to Contested if current U.S. discontinuities spread to core security obligations or if matched PRC commitments show comparable specificity, authorization, resources, and trigger fulfillment.

up to us partner lead

A lead requires repeated, timely fulfillment across core and non-security portfolios plus repair of current stewardship and continuity failures.

to prc centered edge

Requires a representative PRC obligation ledger showing higher fulfillment and government-change persistence under common definitions, not simply greater central policy stability.

current call robust to
  • Treating historical alliance reliability as calibration rather than U.S. performance
  • Giving full positive credit to PRC development delivery and the active DPRK treaty
  • Counting all current U.S. withdrawals as one correlated political shock
current call not robust to
  • Undisclosed PRC trigger and fulfillment records
  • A core U.S. alliance abandonment or funding failure
  • A common program-level commitment universe that reverses observed conversion rates

Fact confidence: High · Comparative confidence: Low · Observability: limited

Return to the position →
05 · Usable accessU.S. / partner edge

RulingThe U.S./partner system has the clearest remaining comparative advantage, but the global category is an edge rather than a lead. It has an order-of-magnitude broader verified network of bases and access sites, supported by treaties, compacts, SOFAs, host investments, persistent presence, prepositioning, and exercises. PRC access is real at Djibouti and increasingly at Ream, and Chinese finance, ownership, and naval familiarity span many ports. Yet most of that wider port network remains conditional potential rather than documented crisis access. Turkey, Niger, Philippine caveats, Cambodian agency, classified terms, and the underobserved nonmilitary portions of the category preclude a global lead.

Strongest counterevidence

  • Turkey blocked combat use in 2003 and Niger forced a U.S. withdrawal in 2024.
  • Philippine officials have imposed public scenario caveats even as EDCA expands.
  • China has an acknowledged Djibouti base, persistent PLAN use at Ream, and naval activity at many Chinese-financed or owned ports.
  • A U.S. ship visit to Ream and Cambodian denials of exclusivity preserve host agency and complicate a simple PRC-base label.
  • Neither side publishes a complete inventory of crisis permissions, capacity, intelligence access, or denied requests.

Why not another call

  • Why not us partner lead — The category extends beyond public military basing; sovereign hosts have denied and revoked U.S. use, national caveats are scenario-specific, and PRC access pathways are expanding faster than formal-base counts imply.
  • Why not contested — The U.S. advantage rests on positive verified chains from agreement through funded facility, persistent presence, exercise, and real activation across many countries, not on missing PRC disclosure.
  • Why not prc centered edge or lead — PRC port finance and naval activity are real but generally do not establish legal, exclusive, timely crisis access or matched capacity.

Rival explanations applied

DC-A01 Strong. Hosts grant access because of shared threats and local incentives, not simply U.S. diplomatic quality.

DC-A02 Strong. Military scale and finance create much of the physical access; this category credits the political and legal conversion only.

DC-A03 Material. Hosts hedge by granting visits, commercial access, or limited permissions to multiple powers.

DC-A05 Strong. Elections, coups, courts, and public opinion can change permissions quickly.

DC-A06 Strong. Classified agreements and operations create asymmetric uncertainty for both systems.

Sensitivity and revision conditions

down to contested

Move to Contested if scenario-specific host permissions remove a large share of U.S. sites or if PRC-owned and financed ports demonstrably provide repeatable crisis access and logistics at representative scale.

up to us partner lead

A global lead requires common scenario tests confirming access, capacity, timing, and survivability across military and nonmilitary functions and showing denial risk is not outcome-changing.

to prc centered edge

Requires a verified network of PRC legal access, host approval, capacity, exercise, and activation that outperforms matched U.S. sites rather than a count of financed ports.

current call robust to
  • Giving full positive credit to Djibouti, Ream, and observed PRC naval activity
  • Excluding commercial ports without crisis permission
  • Retaining Turkey and Niger as decisive U.S. countercases
current call not robust to
  • Classified agreements that materially expand PRC access
  • A major host withdrawal across Japan, ROK, Philippines, or Europe
  • A broader definition that treats commercial finance as usable access without a legal-permission gate

Fact confidence: High · Comparative confidence: Moderate · Observability: partial

Return to the position →
06 · Coalition activationU.S. / partner edge

RulingThe U.S./partner system receives a provisional edge because NATO converted a high-stakes invasion trigger into same-day political authorization, delegated command authority, and a first collective-defence deployment within four to five days. The available SCO stress cases produced declaratory coordination, a directly observed consensus failure, and later formal repair, but not a comparable executable collective response. The edge is narrow and low-confidence: NATO's Afghanistan endgame is a major countercase, both the World Bank and AIIB activated material COVID-19 facilities rapidly, the SCO is not a collective-defence alliance, and the matched event universe remains small.

Strongest counterevidence

  • NATO's own Afghanistan lessons record delayed reporting, constrained political direction, insufficient allied discussion of the U.S.-Taliban agreement, and an evacuation outside NATO command and control.
  • AIIB moved from early preparation to a launched COVID-19 facility and first Board approvals within weeks, showing genuine PRC-initiated multilateral activation.
  • The SCO repaired the June 2025 ministerial consensus failure at the September heads-of-state summit, so one failed statement does not establish durable paralysis.
  • NATO's 2022 response benefited from direct threat, years of warning, standing forces, and a collective-defence mandate unavailable in the PRC-centered comparison.

Why not another call

  • Why not us partner lead — The event sample is too small and heterogeneous; NATO's Afghanistan endgame is a serious countercase; World Bank-AIIB pandemic activation is competitive; and member-level participation denominators remain incomplete.
  • Why not contested — The U.S./partner system has repeatedly observed political-to-command machinery for a live high-stakes security trigger, while the current PRC-centered security sample shows declarations, coordination, and repair but no comparable collective authorization-to-action chain.
  • Why not prc centered edge or lead — AIIB is strong positive counterevidence in a public-goods crisis, but no matched portfolio shows PRC-centered mechanisms activating executable collective responses more reliably across the declared category universe.

Rival explanations applied

DC-A01 Strong. Direct European threat, prior warning, geography, and aligned national interests explain part of NATO's speed and reduce causal credit to institutional machinery.

DC-A02 Strong. NATO's command capacity and members' military resources are material enablers; this domain credits only the conversion from political decision to coordinated authorization and role assignment.

DC-A03 Strong. NATO, the SCO, the World Bank, and AIIB have different mandates. The ruling does not penalize an institution for declining tasks outside its remit, but system-level capability still records whether an executable mechanism exists.

DC-A04 Material. Public records reveal formal decisions more readily than private consultation and bilateral activation, especially in the PRC-centered system.

Sensitivity and revision conditions

down to contested

Move to Contested if a mandate-adjusted sample finds PRC-centered mechanisms matching U.S./partner trigger-to-authorization and participation performance, if additional NATO or U.S.-anchored stress cases repeat the Afghanistan consultation failure, or if national-interest controls remove most institutional conversion credit.

up to us partner lead

A lead requires a larger predeclared event universe with common clocks, eligible-role denominators, national caveats, and repeated U.S./partner advantage across security and non-security activation after controlling for mandate and threat.

to prc centered edge

A PRC-centered edge requires repeated real-event cases in which PRC-centered mechanisms authorize and coordinate consequential assigned action faster or more completely than matched U.S./partner mechanisms.

current call robust to
  • Excluding statements, exercises without after-action evidence, membership totals, and institution charts from decisive activation evidence.
  • Treating AIIB and World Bank output as multilateral institutional action rather than unilateral PRC or U.S. action.
  • Counting the SCO's September 2025 declaration as successful formal repair rather than persistent paralysis.
current call not robust to
  • A rule that prohibits any system-level position until institutions with identical mandates and threat environments can be matched.
  • Undisclosed bilateral or party-state activation evidence that materially changes the observed PRC-centered event portfolio.
  • A larger representative sample showing NATO's 2022 speed is an outlier rather than repeatable capability.

Fact confidence: High · Comparative confidence: Low · Observability: limited

Return to the position →
07 · Coalition deliveryU.S. / partner edge

RulingThe U.S./partner system receives a provisional edge because it shows a broader independently observed record of consequential, high-cost delivery and repair: NATO moved from authorization to deployed force, the World Bank scaled a rapid pandemic portfolio subject to independent evaluation, G7 members implemented a substantial selected commitment set, and European partners replaced much of the 2025 U.S. Ukraine-aid withdrawal. It is not a lead. AIIB produced a large and partly completed COVID-19 portfolio, the state-affiliated BRICS study reports implementation comparable in headline rate to the G7 sample, burden sharing inside the U.S./partner system is highly uneven, and many delivery totals lack common acceptance and outcome standards. The call rests on breadth and critical-role evidence, not on lower PRC-centered observability.

Strongest counterevidence

  • AIIB approved 68 COVID-19 operations totaling $18.48 billion across 26 members, and its internal review reported strong completion among the 11 projects with completion notes.
  • The RANEPA BRICS report records an 83 percent average selected-commitment compliance rate, slightly above the non-equivalent G7 sample's 80 percent headline rate.
  • The United States scored below the G7 average in the selected 2025 portfolio, so autonomous partner implementation is not evidence of unilateral U.S. delivery.
  • European Ukraine support in 2025 did not fully replace every U.S. shortfall, and contributions were highly concentrated among northern European donors.
  • NATO's Afghanistan endgame shows that sustained inputs and activation can coexist with failed consultation, coordination, bargaining, and durable effect.

Why not another call

  • Why not us partner lead — AIIB and BRICS evidence refutes an empty-delivery thesis, U.S./partner burden sharing is uneven, and common accepted-output, critical-role, and outcome denominators remain incomplete.
  • Why not contested — Even after crediting AIIB and BRICS implementation, the U.S./partner portfolio contains a wider independently observed range of high-cost, time-sensitive, and critical-role delivery plus demonstrated partner substitution under stress.
  • Why not prc centered edge or lead — The positive AIIB and BRICS cases do not yet form a comparably broad, independently verified portfolio of critical-role delivery across security, public goods, and economic coordination.

Rival explanations applied

DC-A01 Strong. Threat and national interests explain much of Ukraine delivery and burden variation; institutional machinery receives credit only where coordination, substitution, or a common vehicle is observed.

DC-A02 Strong. Wealth, inventories, military capacity, and development-bank capital enable outputs. The category records political conversion and coalition role completion rather than re-crediting material capacity.

DC-A03 Strong. Disclosure asymmetry favors Western-system observation. The ruling therefore credits positive AIIB completion evidence and does not convert missing PRC-centered data into a deficit.

DC-A04 Material. World Bank age and scale explain part of its portfolio advantage, while AIIB's younger age and co-financing model complicate causal attribution.

Sensitivity and revision conditions

down to contested

Move to Contested if independent project and obligation ledgers show AIIB, BRICS, SCO, or FOCAC matching U.S./partner breadth and critical-role completion, if a common-output reanalysis removes the Ukraine and World Bank advantages, or if partner substitution proves episodic rather than repeatable.

up to us partner lead

A lead requires repeated accepted-output and critical-role completion across a predeclared portfolio, balanced burden or reliable substitutes, and no material PRC-centered portfolio approaching the same breadth after disclosure adjustment.

to prc centered edge

A PRC-centered edge requires independently verified, mandate-adjusted delivery across multiple institutional families, including critical roles, timeliness, quality, repair, and participant retention, not approvals or selected compliance alone.

current call robust to
  • Crediting AIIB's reported reviewed-project completion and BRICS selected-commitment implementation as genuine positive counterevidence.
  • Removing headline pledges, raw finance, membership counts, and unaccepted commitments from decisive evidence.
  • Crediting European 2025 Ukraine aid to participating partners rather than automatically to the United States.
current call not robust to
  • A rule requiring identical mandates and outputs across every matched case before any portfolio-level position.
  • New independent PRC-centered completion data that closes the current breadth and critical-role gap.
  • Evidence that tracked allocations systematically overstate timely, usable delivery or that institutional coordination did not contribute to partner substitution.

Fact confidence: High · Comparative confidence: Low · Observability: limited

Return to the position →
08 · Institutional leadershipU.S. / partner edge

RulingThe U.S.-anchored system retains an edge because NATO, the G7, the OECD, and the World Bank demonstrate deeper and more routinely monitored conversion from collective decisions into implementation than the available BRICS, SCO, FOCAC, and AIIB portfolio. It is not a lead: current U.S. arrears, lower-than-group G7 compliance, WTO Appellate Body obstruction, and withdrawals from WHO, the Paris Agreement, and other bodies materially weaken stewardship; PRC-centered institutions also show real attraction, financing, oversight, and partial implementation. The finding credits participating partners separately and does not treat institutional output as unilateral U.S. or PRC output.

Strongest counterevidence

  • The U.S. position was the larger coalition in major UN votes on nonproliferation, fissile-material negotiations, and responsible space behavior.
  • The G7's selected 2025 commitments averaged 80 percent compliance even though the United States scored 73 percent, showing autonomous partner-system resilience.
  • AIIB has broad membership, large commitments, Board-reporting oversight, and published project learning, refuting an empty-institution thesis.
  • BRICS compliance research reports material implementation, but its selection and institutional provenance limit the strength of the inference.
  • Both the United States and China carried substantial UN arrears at the April 2026 cutoff and both have recently used Security Council vetoes.
  • SCO expansion coexists with competing leadership and blocked PRC initiatives, while its narrow mandate may make alliance-style tests inappropriate.

Why not another call

  • U.s. / partner lead — Rejected because current U.S. behavior weakens institutions it helped build, the PRC position aligns with much broader UN General Assembly coalitions on many recorded votes, and AIIB plus expanded BRICS demonstrate material alternative institution-building capacity.
  • Contested — Rejected for the global category because the Western basket still shows a broader set of mature, repeatable implementation, peer-review, evaluation, and political-to-operational mechanisms across consequential institutions. The PRC-centered basket has not yet matched that full-chain depth with comparably independent evidence.
  • Prc-centered edge — Rejected because vote convergence and membership growth do not establish PRC sponsorship or implementation, BRICS delivery evidence is selectively coded and partly state-affiliated, SCO leadership is shared and contested, and AIIB is a professional multilateral institution whose output cannot be assigned wholesale to Beijing.
  • U.s. / partner lead recovery gate — Would require material repair of current stewardship failures and sustained implementation, not simply greater inherited institutional scale.

Rival explanations applied

DC-A01 Material. Threat and interest alignment explains part of NATO and UN vote behavior; it does not erase the institutional conversion machinery but reduces causal credit to either leading state.

DC-A02 Material. Wealth, military protection, market access, and finance enable institutional output; only the political conversion mechanism is credited here.

DC-A03 Strong in the BRICS, SCO, and AIIB basket. Multi-membership and hedging make attraction different from alignment.

DC-A04 Strong. Institutional age and mandate explain substantial Western execution depth, which is why the position is an edge with moderate confidence rather than a lead.

DC-A05 Material. U.S. administration change explains current stewardship discontinuity and must remain distinct from the durability of partner institutions.

DC-A06 Strong. Western institutions disclose more failures and independent evaluations; PRC-centered evidence relies more heavily on official and state-affiliated sources, limiting symmetric confidence.

Sensitivity and revision conditions

down to contested

The global call falls to Contested if U.S. withdrawals, arrears, and WTO paralysis persist through the next annual checkpoint while PRC-centered institutions show independently verified implementation and retention across at least two distinct institutional families, or if matched attribution removes most Western-system credit.

up to us partner lead

A lead requires sustained NATO and G7 conversion plus material U.S. repair of UN funding and WTO adjudication and reengagement or durable substitution for major withdrawals, while PRC-centered portfolios continue to lack independently verified delivery.

to prc centered edge

A PRC edge requires more than UN vote convergence: predeclared proposal sponsorship, independently verified BRICS, SCO, FOCAC, and AIIB implementation, participant retention under stress, and evidence that current Western institutional failures reduce delivered outcomes.

current call robust to
  • Removing raw membership, voting-power, post-count, meeting-count, and finance-size inputs from the decisive evidence.
  • Treating all 2025-2026 U.S. withdrawals as one correlated policy shock rather than independent observations.
  • Treating the RANEPA BRICS report as state-affiliated evidence and excluding it from independent corroboration.
  • Crediting AIIB and World Bank outputs to the institutions and participating members rather than wholesale to Beijing or Washington.
current call not robust to
  • A decision rule that values current stewardship over inherited implementation depth as the sole decisive function.
  • New matched evidence showing the apparent Western implementation advantage is mainly institutional age, mandate, or case selection rather than repeatable system capability.
  • Independent obligation-level evidence showing PRC-centered delivery and repair equal to or greater than the Western portfolio.

Fact confidence: High · Comparative confidence: Moderate · Observability: partial

Return to the position →
09 · Bargaining outcomesInsufficient comparative basis

RulingInsufficient comparative basis is the completed judgment. Both systems have real positive cases and serious failures: the U.S.-anchored P5+1 obtained initial verified JCPOA implementation before U.S. withdrawal; Beijing facilitated Saudi-Iran diplomatic restoration after earlier Iraqi and Omani work; the Gulf coalition achieved Kuwait's liberation; the Defeat ISIS coalition ended territorial control; and Afghanistan failed to produce a durable political settlement. These cases do not form a valid comparative portfolio. Objectives, opponents, instruments, issue stakes, compliance windows, and coalition counterfactuals differ, and military or economic power often supplies the proximate causal mechanism. Calling the category Contested would manufacture a tie from incomparable evidence.

Strongest counterevidence

  • The P5+1 and IAEA record demonstrates that a coalition can obtain and verify a complex negotiated objective.
  • Saudi-Iran embassy restoration demonstrates a concrete PRC-facilitated diplomatic result.
  • Afghanistan and the later JCPOA trajectory demonstrate that initial agreement or enormous coalition effort does not guarantee durable effect.
  • Gulf and ISIS cases demonstrate outcomes but rely principally on military and economic mechanisms outside this domain.

Why not another call

  • Why not contested — Contested requires positive evidence for both systems under a common rule; this portfolio contains unlike cases, not a matched tie.
  • Why not us partner edge or lead — U.S. cases combine genuine successes with self-inflicted durability failure and major cross-domain confounding, without a matched PRC case set.
  • Why not prc centered edge or lead — Saudi-Iran restoration is important but is one bounded mediation case, not evidence of superior coalition bargaining across issues.

Rival explanations applied

DC-A01 Decisive. Opponent and partner interests often explain agreement or defection.

DC-A02 Decisive. Sanctions, force, market access, and security guarantees frequently dominate bargaining leverage.

DC-A05 Strong. Government change altered JCPOA durability and many counterpart compliance choices.

DC-A06 Strong. Private demands, side payments, abandoned objectives, and failed initiatives are systematically underobserved.

Sensitivity and revision conditions

minimum portfolio

At least three predeclared matched pairs in two issue families, with one success and one failure eligible on each side.

required fields
  • ex ante objectives
  • baseline
  • final terms
  • decisive concessions
  • 12- and 24-month compliance
  • coalition-only mechanism
  • rival explanations
  • sender and partner costs
current shortfall

No matched pair satisfies all required fields; public and private objectives plus counterfactuals remain incomplete.

Fact confidence: High · Comparative confidence: Not assessed · Observability: limited

Return to the position →
10 · Coalition durabilityU.S. / partner edge

RulingThe U.S./partner system has a low-confidence durability edge. Its core alliances have survived decades of government turnover, adapted after strategic shocks, retained or added capable members, and repeatedly repaired internal disputes while continuing costly role delivery. PRC-centered forums are also durable: BRICS has expanded despite member disputes, FOCAC has persisted for twenty-five years, and the SCO has repaired consensus failures. What separates the systems is the depth of observed costly-role persistence and substitution, not mere institutional age. The edge is sharply narrowed by current U.S. withdrawals and aid disruption, transatlantic political strain, host-access reversals, and the fact that shared threat and autonomous partner action explain much U.S.-system resilience.

Strongest counterevidence

  • BRICS retained members through major disputes and expanded to a larger member and partner network.
  • FOCAC has sustained a regular China-Africa architecture for twenty-five years.
  • The SCO repaired a public consensus failure rather than fracturing.
  • U.S. withdrawals, aid disruption, host revocations, and transatlantic uncertainty show substantial current fragility.
  • European substitution for U.S. Ukraine support demonstrates partner-system durability but cannot be credited automatically to Washington.

Why not another call

  • Why not us partner lead — Current U.S. government-change discontinuity, host revocations, burden concentration, and partner-led substitution are too material; PRC-centered institutions show genuine expansion and persistence.
  • Why not contested — U.S.-anchored alliances have a broader positive record of continued costly-role delivery, adaptation, and repair under severe shocks, not merely older charters or more members.
  • Why not prc centered edge or lead — BRICS, FOCAC, and SCO longevity and expansion are strong, but public evidence of persistent costly-role fulfillment and substitution under pressure is thinner and often declaratory.

Rival explanations applied

DC-A01 Strong. Russian and Chinese pressure and local security interests explain much allied persistence.

DC-A02 Strong. U.S. and European material power makes role substitution possible; finance and trade sustain PRC forums.

DC-A03 Strong. BRICS and SCO retention may reflect low exit costs and hedging rather than cohesive alignment.

DC-A04 Strong. NATO's age, mandate, and standing machinery make it unlike younger and looser forums.

DC-A05 Decisive for current confidence. Elections and coups explain U.S. continuity shocks and host-access changes.

DC-A06 Material. Open systems disclose dispute and shortfall data more fully than PRC-centered institutions.

Sensitivity and revision conditions

down to contested

Move to Contested if current U.S. discontinuities reduce core allied role delivery or if PRC-centered institutions document comparable costly-role retention and repair across multiple shocks.

up to us partner lead

A lead requires sustained U.S. continuity plus representative role-level 12- and 24-month evidence across regions and issues, with reduced single-provider dependence.

to prc centered edge

Requires PRC-centered coalitions to retain and replace costly roles through leadership turnover, coercion, and member disputes at higher rates under common denominators.

current call robust to
  • Crediting European substitution to partners rather than the United States
  • Treating BRICS and FOCAC longevity as strong positive evidence
  • Counting 2025-2026 U.S. withdrawals as a decisive countercase
current call not robust to
  • A major NATO or Indo-Pacific alliance defection
  • Independent PRC-centered role-level durability data
  • A methodology that treats membership retention as equivalent to material role retention

Fact confidence: High · Comparative confidence: Low · Observability: limited

Return to the position →
Falsifiable hypotheses

The synthesis is conditional.

These hypotheses preserve the strongest alternative, disconfirming evidence, open gaps, and the present adjudication.

DC-H01The U.S./partner system converts commitments into coordinated action more reliably when a standing alliance or institution supplies decision and implementation machinery.

Strongest alternative

Threat intensity, U.S. material capacity, or case selection may explain participation and delivery rather than standing diplomatic machinery.

What would disconfirm it

Matched cases in which standing U.S.-anchored machinery repeatedly fails to authorize, activate, or deliver while a comparable PRC-centered mechanism succeeds under similar stakes and partner incentives.

Open gaps

The applied ledger contains 19 observations across security activation, pandemic development-bank response, selected club implementation, and Ukraine delivery, but mandate-adjusted event breadth, eligible-role denominators, accepted outputs, and 24-month durability remain incomplete.

Required tests

  • Separate formal decision speed from lowest-common-denominator output.
  • Control for direct threat, geography, U.S. military or economic contribution, and preexisting partner preference.
  • Retain opt-outs, caveats, late delivery, abandoned commitments, and post-event erosion.
DC-H02The PRC has stronger formal and party-to-party reach in parts of Asia and the Global South than embassy counts alone reveal.

Strongest alternative

Reported party and forum relationships may be ceremonial, dormant, or driven by economic exposure and elite incentives rather than usable diplomatic access.

What would disconfirm it

Entity-level sampling shows low recency or reciprocity, weak access across government changes, or no higher conversion into agenda entry, warning, coordination, or policy than matched U.S. networks.

Open gaps

The PRC network is documented mainly by actor claims; a matched entity-level dataset for U.S. plural and PRC party-state channels does not yet exist.

Required tests

  • Verify reported relationships at the partner level and distinguish governing, opposition, and ceremonial contacts.
  • Compare network recency, reciprocity, access, and policy conversion rather than totals.
  • Separate PRC party-state tasking advantages from legitimacy, autonomy, and durability costs.
DC-H03U.S. institutional leadership is strongest inside Western security and economic institutions but weaker when current stewardship, representativeness, and global legitimacy are included.

Strongest alternative

The United States may remain the indispensable provider and agenda setter even when it challenges an institution, while survey criticism and governance inequities do not show PRC leadership.

What would disconfirm it

Across predeclared Western and universal-institution portfolios, U.S.-sponsored proposals fail, U.S. obligations are not implemented, institutions remain unrepaired, or members shift to alternatives at rates inconsistent with a U.S. edge.

Open gaps

The first adjudication now applies 38 observations across Western, universal, and PRC-centered baskets, but proposal sponsorship, private bargaining, regionally matched implementation, and independent obligation-level PRC-centered delivery records remain incomplete.

Required tests

  • Assess Western, universal, and PRC-centered institutional baskets separately.
  • Distinguish governance power from agenda adoption, implementation, stewardship, inclusion, and legitimacy.
  • Apply stated institutional rules to both powers and include counterexamples to each strongest claim.
DC-H04PRC-centered forums broaden participation and agenda reach, while consensus rules and member heterogeneity may constrain activation and delivery.

Strongest alternative

Flexible consensus and heterogeneous membership may be deliberate strengths that preserve partner agency and enable selective, durable cooperation without alliance-like activation.

What would disconfirm it

Obligation-level records show high authorization, resource, implementation, and retention rates across consequential PRC-centered initiatives despite divergent member preferences.

Open gaps

The SCO stress cases and AIIB and BRICS implementation evidence now test the claim directly, but independent obligation inventories and representative activation and delivery samples for BRICS, SCO, and FOCAC remain incomplete.

Required tests

  • Evaluate each forum against its stated function rather than an alliance template.
  • Separate new members, partners, observers, meeting participants, and material implementers.
  • Trace named commitments through owners, resources, milestones, delivered output, and participant retention.
DC-H05U.S. plural informal networks may be broad and durable but less centrally taskable; PRC party-state networks may be more taskable but more exposed to legitimacy and partner-autonomy costs.

Strongest alternative

Central taskability and plural durability are assumed organizational traits that may reverse by country, sponsor, issue, and political regime.

What would disconfirm it

Matched network cohorts show no systematic difference in taskability, reciprocity, policy conversion, autonomy costs, or persistence after leadership change.

Open gaps

Private networks are intrinsically underobserved and the available evaluations use different participant, outcome, and disclosure standards.

Required tests

  • Code channel sponsor and form without assuming public, private, party, or civil actors act as a unified national system.
  • Require an observed conversion from relationship to consequential access, warning, coordination, or policy.
  • Measure partner-defined benefits, reciprocal influence, attrition, and behavior after political turnover.
DC-H06Both systems will look less capable when commitments are followed through authorization, resources, usable access, activation, delivery, outcomes, and durability rather than counted at signature.

Strongest alternative

Some diplomatic commitments are intentionally expressive, deterrent, or constitutive; demanding downstream material output from every instrument could create a false failure rate.

What would disconfirm it

Preclassified commitments show little attrition at the stages applicable to their stated form and purpose, or apparent attrition is explained by incorrect denominators rather than nonperformance.

Open gaps

Complete, comparable commitment universes with implementation deadlines are not yet built for either system.

Required tests

  • Classify each commitment's intended function before selecting applicable causal stages.
  • Use the same obligation specificity and milestone rules for both systems.
  • Do not penalize a nonmaterial political commitment for lacking resources unless its own objective requires them.
DC-H07Bargaining outcomes are more case- and issue-specific than the other capability positions and may remain unscored longest.

Strongest alternative

A sufficiently broad portfolio may reveal systematic differences in coalition legitimacy, leverage, information, or enforcement that generalize across cases.

What would disconfirm it

Predeclared matched cases show a stable cross-issue conversion advantage after controlling for material leverage, objective difficulty, threat, and partner preferences.

Open gaps

Objectives, private concessions, side payments, and credible counterfactuals are frequently unobservable or selected after the fact.

Required tests

  • Recover objectives documented before the outcome and retain abandoned or revised demands.
  • Test opponent compliance and reversal at 12 and 24 months.
  • State and challenge the counterfactual contribution of the coalition apart from military and economic power.
Mandatory audit standard

Sixteen tests applied before publication.

  1. The U.S./partner and PRC/partner observations use the same unit, period, geography, issue, denominator, and stage wherever a direct comparison is claimed.
  2. A deficit, failure, or missing observation on one side never establishes an advantage for the other.
  3. Insufficient comparative basis is used instead of Contested when positive evidence for both sides is absent.
  4. Partner capacity, access, votes, and delivery are credited only when observed, authorized, relevant, and timely.
  5. Formal posts, visits, dialogue counts, signatures, memberships, and favorable sentiment remain inputs unless a defined conversion is observed.
  6. First-party material establishes rules, commitments, contributions, and actor claims but not comparative effectiveness, legitimacy, or causal attribution by itself.
  7. Sources from the same government, institution, or dataset family are not counted as independent corroboration.
  8. Formal and informal channels remain distinct, while hybrid party-state and quasi-official channels retain their actual sponsor and legal character.
  9. Western, universal, and PRC-centered institutions are assessed in separate baskets before any cross-institution synthesis.
  10. U.S. liberal-order leadership is tested through compliance, stewardship, repair, inclusion, and implementation, including serious counterexamples.
  11. PRC alternative-institution leadership is tested through partner agency, implementation, additionality, inclusion, and durability, including serious counterexamples.
  12. Every global position survives regional and issue sensitivity checks or is narrowed to the slices the evidence supports.
  13. Every outcome claim declares the objective, baseline, endpoint, compliance window, rival explanations, and coalition counterfactual.
  14. Successes, failures, mixed outcomes, ongoing cases, abandoned initiatives, opt-outs, and non-events are eligible for the sample.
  15. Recent cases are not given durability credit before 12- and 24-month checkpoints.
  16. No weighted overall diplomatic-power composite is calculated.