Behavior change beyond intervention: an activity-theoretical perspective on human-centered design of personal health technology

IntroductionModern personal technologies, such as smartphone apps with artificial intelligence (AI) capabilities, have a significant potential for helping people make necessary changes in their behavior

A data-centric perspective on designing AI foundation models for healthcare

Post Content

Disclosure in the era of generative artificial intelligence

Generative artificial intelligence (AI) has rapidly become embedded in academic writing, assisting with tasks ranging from language editing to drafting text and producing evidence. Despite

Longitudinal assessment of chorea in Huntington’s disease using digital passive monitoring

npj Digital Medicine, Published online: 25 April 2026; doi:10.1038/s41746-026-02661-y Longitudinal assessment of chorea in Huntington’s disease using digital passive monitoring

Connections across regional glymphatic clearance, neural activity and amyloid-β deposition in cortex

Neural activity inevitably produces waste, which promotes neurodegeneration with topographic features. The glymphatic system is important for waste clearance. However, the spatial characteristics of glymphatic

Soft-Label Governance for Distributional Safety in Multi-Agent Systems

April 23, 2026

arXiv:2604.19752v1 Announce Type: cross
Abstract: Multi-agent AI systems exhibit emergent risks that no single agent produces in isolation. Existing safety frameworks rely on binary classifications of agent behavior, discarding the uncertainty inherent in proxy-based evaluation. We introduce SWARM (textbfSystem-textbfWide textbfAssessment of textbfRisk in textbfMulti-agent systems), a simulation framework that replaces binary good/bad labels with emphsoft probabilistic labels $p = P(v=+1) in [0,1]$, enabling continuous-valued payoff computation, toxicity measurement, and governance intervention. SWARM implements a modular governance engine with configurable levers (transaction taxes, circuit breakers, reputation decay, and random audits) and quantifies their effects through probabilistic metrics including expected toxicity $mathbbE[1-p mid textaccepted]$ and quality gap $mathbbE[p mid textaccepted] – mathbbE[p mid textrejected]$. Across seven scenarios with five-seed replication, strict governance reduces welfare by over 40% without improving safety. In parallel, aggressively internalizing system externalities collapses total welfare from a baseline of $+262$ down to $-67$, while toxicity remains invariant. Circuit breakers require careful calibration; overly restrictive thresholds severely diminish system value, whereas an optimal threshold balances moderate welfare with minimized toxicity. Companion experiments show soft metrics detect proxy gaming by self-optimizing agents passing conventional binary evaluations. This basic governance layer applies to live LLM-backed agents (Concordia entities, Claude, GPT-4o Mini) without modification. Results show distributional safety requires emphcontinuous risk metrics and governance lever calibration involves quantifiable safety-welfare tradeoffs. Source code and project resources are publicly available at https://www.swarm-ai.org/.

Subscribe for Updates

Copyright 2025 dijee Intelligence Ltd. dijee Intelligence Ltd. is a private limited company registered in England and Wales at Media House, Sopers Road, Cuffley, Hertfordshire, EN6 4RY, UK registration number 16808844