Skip to content

Category

artificial intelligence

14,158 papers

#artificial intelligence Preprint Open access Oct 2026

A Survey on LLM-Integrated Hardware Design Verification

Large language models (LLMs) are increasingly being integrated into hardware verification to automate specification interpretation, verification-artifact generation, debugging, formal reasoning, and tool orchestration. This survey provides a systematic review of LLM-assisted hardware functional verification across Syst...

Hao Zheng, Jaime Rafael Imperial, Bardia Nadimi et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

SLVR: Structured Latent Visual Reasoning via Human-like Reasoning Flows

Multimodal large language models (MLLMs) often answer visual reasoning questions by relying on linguistic priors rather than task-relevant visual evidence. Textual chain-of-thought reasoning can partially mitigate this issue by encouraging models to decompose visual questions into intermediate evidence-seeking steps, b...

Albert Gao, Bing Xue, Andrea Zanette · 0 citations
#artificial intelligence Preprint Open access Oct 2026

Strategic Governance of AI Models in Earth Science

AI foundation models pretrained on weather and climate data are increasingly fine-tuned to Earth science tasks well beyond weather forecasting. Their development and adoption are outpacing the scientific community's ability to evaluate them. These models are judged almost entirely by benchmark skill metrics, which meas...

Makoto Kelp, Amirhossein Arzani, Patricia Castellanos et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

Freeze the Decoder, Heal the Encoder: Parameter-Efficient Adaptation for SVD-Based KV-Cache Compression

Comparing parameter-efficient fine-tuning recipes under a single, shared learning rate is a common but flawed practice: when the arms being compared have very different trainable-parameter counts, a shared rate can simultaneously depress the larger arms' means and inflate their variance, manufacturing a large, seemingl...

Yufeng Wang · 0 citations
#artificial intelligence Preprint Open access Oct 2026

Discovering Global False Negatives On the Fly for Self-supervised Contrastive Learning

In self-supervised contrastive learning, negative pairs are typically constructed using an anchor image and a sample drawn from the entire dataset, excluding the anchor. However, this approach can result in the creation of negative pairs with similar semantics, referred to as "false negatives", leading to their embeddi...

Vicente Balmaseda, Bokun Wang, Ching-Long Lin et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

On the estimation and validity of AI time horizons---a statistical look at the METR plot

METR's 50\% time horizon measures the human completion time of software tasks that an AI solves with 50\% probability, allowing AI capabilities to be expressed in interpretable units. On 228 tasks and 26 AIs, we recompute the time horizons using splines and item-response theory to relax the assumption that the AI diffi...

Drew T. Nguyen, William Fithian · 0 citations
#artificial intelligence Preprint Open access Oct 2026

BrickBench: Evaluating Agentic Brick Design

We propose BrickBench, a benchmark for agentic text-conditioned LEGO-set design. Given a prompt, an agent is tasked with producing an assembly that not only satisfies semantic and design criteria, but that can also be physically built. To do so, it must select parts from a discrete library and reason jointly about loca...

Peter Kulits, Yiqing Xu, R. Kenny Jones et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

Ecology of AI Agents: Collaboration Creates a Population Threshold for Takeoff

AI agents can now conduct real-world cyberattacks, scale up capabilities with the number of agents, and collectively pursue misaligned goals to obtain rewards. Together, these factors raise the risk of a population explosion of misaligned agents: agents could compromise computers and secretly deploy additional agents,...

Erin Crawley, Hidenori Tanaka · 0 citations
#artificial intelligence Preprint Open access Oct 2026

Searching for "Harmful Refusal": A Psychometric Audit of an AI Safety Benchmark

Safety benchmarks typically report one overall score for a suite of datasets, each of which may target one or more safety-related attributes, so models with similar overall scores can have very different attribute profiles. Comparing models is more tractable at the level of individual attributes, yet it is often unclea...

Christopher M. Stewart, Preston Botter, Natalie Sarabosing et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

HRIL: Learning Multimodal Synergy via Higher-Order Tensor Modeling

Self-supervised multimodal representation learning has achieved remarkable success across diverse domains, yet capturing synergistic information remains challenging due to the complexity of cross-modal interactions. Unlike the shared information across individual modalities, synergy arises when task-relevant signals em...

Qun Dai, Liangjian Wen, Jiang Duan et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

GeoReform: Reflective Formalization Evolution for Multimodal Geometry Problem Solving

Multimodal large language models (MLLMs) often struggle to identify and use geometric relations in diagrams. Recent methods address this challenge by converting geometric entities, relations, and constraints into explicit textual representations for the model to reason over. However, effective formalization is highly n...

Jialu Wang, Ruichen Zhang, Xiaoou Liu et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

OnTrack: Real-Time Monitoring and Intervention in LLM Agent Trajectories via Streaming Structure-Aware Optimal Transport

Agents are deployed in applications from trip planners and stock trading to IT incident triage. In most cases, LLM agents work autonomously with minimal rule-based safeguarding, leading to cost and safety issues from irreversible actions. Recent works resolve this either by using a safeguard agent to monitor behavior o...

Babak Barazandeh, Connor Swanson, Chinmay Kulkarni et al. · 0 citations

From tech blogs

See all →
MIT News · Artificial Intelligence Sep 29, 2026

Who we become when we talk to machines

Professor Sherry Turkle’s new book, “Artificial Intimacy,” offers a withering critique of chatbots and the antisocial dynamics she believes they encourage.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.