Full-duplex speech models are trained to converse with a person, but they are increasingly made to converse with each other, in self-play data generation, agent societies, and model-based evaluation. In that loop no human absorbs a timing error: each model's turn-taking is the other's input. We ask what timing the loop...
Li-Chen Zhu, Yueqian Lin, Yi-Heng Wang et al.· 0 citations
Personal LLM assistants (health companions, elder-care agents, accessibility aides) are judged by what they remember about a person: a medication or an allergy mentioned in passing and needed days later, so an eviction policy must decide what the cache forgets.
Li-Chen Zhu, Yueqian Lin, Yi-Heng Wang et al.· Proceedings of the 4th Inter...· 0 citations
MUSE (Multimodal Unified Safety Evaluation), an open-source, browser-based, run-centric platform for multimodal safety evaluation, demonstrates the value of run-centric, fine-grained evaluation for characterizing multimodal safety behavior beyond a single binary success metric.
This survey provides a comprehensive overview of techniques that enable GenAI deployment at the edge, covering software optimizations, hardware innovations, and system-level frameworks, with particular emphasis on hardware-focused approaches.
Mozhgan Navardi, Yuzhe Fu, Yueqian Lin et al.· ACM Computing Surveys· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.