Scientific artificial intelligence (AI), spanning foundation models (FMs) to federated data-analysis pipelines, is becoming shared infrastructure across national laboratories, universities, hospitals, and industrial partners. This collaboration creates privacy risks whose natural unit is often an institution's particip...
O. Kotevska, Sumit Kumar Jha, A. Bellet et al.· 1 citation
Multi-agent LLM systems can improve reasoning by pooling diverse perspectives, but their effectiveness depends on coordinating communication, particularly in hidden-profile settings where each agent holds only part of the evidence required for a correct decision. Existing protocols, including fixed schedules, round-rob...
Abhijith Babu, Ramneet Kaur, Vishal Pramanik et al.· 0 citations
NSF-CoT is presented, a neuro-symbolic formal verification method that checks CoT faithfulness step by step for contextual question answering and consistently outperforms causal mediation, perturbation probes, and behavioral monitoring.
Vishal Pramanik, Maisha Maliha, Nathaniel D. Bastian et al.· Annual Meeting of the Associ...· 1 citation
Verifiable Latent Alignments (VLA), an activation-aware framework for monitoring and steering these private communication channels, is introduced and shows that the evaluated private channel attacks can be monitored without training the primary monitor on attack examples and mitigated when matched counterfactual access...
Ramneet Kaur, Pradyumna Chari, Ramesh Raskar et al.· 1 citation
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.