Skip to content

Author

Franck Dernoncourt

We have 11 of 96 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#artificial intelligence Preprint Sep 2026

FlexRouter: Learning Complementary Model Sets for Flexible LLM Routing

Existing Large Language Model (LLM) routing methods score LLMs independently to select top-$k$ models. However, this ignores model correlations and enforces a rigid computational budget. Consequently, routers often select redundant models that share failure modes, limiting the overall probability of success. To address...

Wang Wei, Harry Yang, Tiankai Yang et al. · 1 citation
#artificial intelligence Preprint Sep 2026

Controlled Decoding Attacks on Black-Box LLMs

Manipulating next-token probabilities during generation can bypass the safety alignment of large language models. Existing approaches, however, rely on access to model weights or numerical token probabilities and therefore do not apply to interfaces that return only sampled text. Reconstructing probabilities from sampl...

Jesson Wang, Shawn Li, Wei Yang et al. · 0 citations
#machine learning Preprint Sep 2026

Online Learning with LLM Experts from Limited Feedback

We study adaptive routing of prompts to large language model (LLM) experts to maximize response quality in an online setting with limited feedback. We formulate it as a bandit problem with $K$ actions that represent experts and $d$ features that encode prompts, over a horizon of $T$ rounds. We propose algorithms that s...

Wei Wang, Soumyabrata Pal, Koyel Mukherjee et al. · 0 citations
#artificial intelligence Preprint Sep 2026

Beyond Top-$k$ Skill Retrieval: Diversity-Aware Skill Routing for LLM Agents

Large language model (LLM) agents increasingly rely on external skills, but routing user requests over large skill registries is difficult because many skills are functionally redundant while complex tasks often require complementary skill sets. Existing skill routers typically rank candidates independently by query re...

Wang Wei, Tiankai Yang, Samyadeep Basu et al. · 2 citations
#machine learning Preprint Sep 2026

CRISP: Cliff-awaRe Input-adaptive Sparse Prefilling with Structural-Mass-Motivated Routing

This work replaces the Jensen-Shannon Divergence routing with C_struct, a structural proxy that measures mass at Vertical-Slash compatible positions and reproduces JSD's routing decisions while eliminating both the pooled matmul and subsequent KL divergence overhead.

H. Nguyen, Chien Van Nguyen, Franck Dernoncourt et al. · 0 citations
#artificial intelligence Preprint Aug 2026

Not Worth Another Token: Marginal Value Estimation for Efficient Deep Research Agents

The results show that pruning effectiveness depends more on where pruning is applied than on the specific scoring rule: early pruning yields the largest end-to-end savings, while later pruning mainly refines the final synthesis context.

Harshitha Kolukuluru, Reshma Ashok, Kirat Arora et al. · 0 citations
Preprint Aug 2026

Unifying Graph Neural Networks Through a Common Layer Equation

A common layer equation is introduced that represents covered architectures through seven components: an update domain, channel set, propagation bank, per-channel message maps, channel-fusion operator, ego/residual map, and update map, which exposes the empirical inverse problem of mapping measurable graph and task pro...

S. Navuluru, Siddhartha Shankar Das, B. Ni et al. · 0 citations
Review Aug 2026

Personalized Auto-Research: Towards a True AI Co-Scientist

This work introduces the problem of personalized auto-research, which conditions every stage of the research process on a representation of the individual researcher, and proposes a general and flexible framework that threads a graph-grounded researcher context through retrieval, hypothesis search, experimentation, wri...

B. Ni, Franck Dernoncourt, Hong-Jie Chen et al. · 0 citations
Jul 2026

GRASP: GRanularity-Aware Search Policy for Agentic RAG

GRASP is introduced, a reinforcement learning (RL) framework for training agents to adaptively coordinate complementary retrieval tools during multi-step reasoning, and it is suggested that learning to coordinate retrieval signals and context granularity is critical for agent's correct reasoning.

Varun Gandhi, Jaewook Lee, Shantanu Todmal et al. · 0 citations
Conference Open access 2026

Octopus: Gated Selective Attention for Memory-Bounded Long-Context Inference in Large Language Models

O CTOPUS is proposed, a framework that confers fixed-memory inference onto pretrained Transform-ers without the information loss of linearization and outperforms state-of-the-art linearized baselines on the GSM8K benchmark, demonstrating that learned sparse retention serves as an effective regular-izer for long-horizon...

C. Nguyen, Ryan A. Rossi, L. Van et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.