AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
19 May 2026

Unleashing LLMs in Bayesian Optimization: Preference-Guided Framework for Scientific Discovery

ApplicationsDGX agent

arXiv:2605.17976v1 Announce Type: new Abstract: Scientific discovery is increasingly constrained by costly experiments and limited resources, underscoring the need for efficient optimization in AI for

Unleashing the Potential of Diffusion Models for End-to-End Autonomous Driving

SafetyDGX agent

arXiv:2602.22801v2 Announce Type: replace-cross Abstract: Diffusion models have become a popular choice for decision-making tasks in robotics, and more recently, are also being considered for solving

Unlocking the Potential of Diffusion Language Models through Template Infilling

ResearchDGX agent

arXiv:2510.13870v3 Announce Type: replace-cross Abstract: Diffusion Language Models (DLMs) have emerged as a promising alternative to Autoregressive Language Models, yet their inference strategies rem


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

UNR-Explainer: Counterfactual Explanations for Unsupervised Node Representation Learning Models

ResearchDGX agent

arXiv:2605.17285v1 Announce Type: cross Abstract: Node representation learning, such as Graph Neural Networks (GNNs), has emerged as a pivotal method in machine learning. The demand for reliable expla

Unveiling Memorization-Generalization Coexistence: A Case Study on Arithmetic Tasks with Label Noise

ApplicationsDGX agent

arXiv:2605.18022v1 Announce Type: cross Abstract: Highly over-parameterized models can simultaneously memorize noisy labels and generalize well, yet how these behaviors coexist remains poorly understo

UVTran: Accurate Hole-Filling Parameterization with Transformers

Model ReleasesDGX agent

arXiv:2605.16306v1 Announce Type: cross Abstract: In industrial design, N-sided hole filling is typically formulated as the construction of a single trimmed B-spline surface by minimizing a fairness e

Validate Your Authority: Benchmarking LLMs on Multi-Label Precedent Treatment Classification

Model ReleasesDGX agent

arXiv:2605.17691v1 Announce Type: cross Abstract: Automating the classification of negative treatment in legal precedent is a critical yet nuanced NLP task where misclassification carries significant

Verify-Gated Completion as Admission Control in a Governed Multi-Agent Runtime: A Bounded Architecture Case Study

Model ReleasesDGX agent

arXiv:2605.17998v1 Announce Type: cross Abstract: As multi-agent systems move from short interactions to tool-using workflows with specialized roles and persistent state, completion becomes a runtime-

VeriHGN: Heterogeneous Graph-Based Congestion Prediction for Chip Layout Verification

ResearchDGX agent

arXiv:2603.11075v2 Announce Type: replace-cross Abstract: As Very Large Scale Integration (VLSI) designs continue to scale in size and complexity, layout verification has become a central challenge in

VGGT-CD: Training-Free Robust Registration for 3D Change Detection

Model ReleasesDGX agent

arXiv:2605.16859v1 Announce Type: cross Abstract: 3D change detection from multi-view images is essential for urban monitoring, disaster assessment, and autonomous driving. However, existing methods p

Virtual Nodes Guided Dynamic Graph Neural Network for Brain Tumor Segmentation with Missing Modalities

ResearchDGX agent

arXiv:2605.16880v1 Announce Type: new Abstract: Multimodal magnetic resonance imaging (MRI) is crucial for brain tumor segmentation, with many methods leveraging its four key modalities to capture com

Virtues of Ordered Chaos: Planning with Topple Actions in Tabletop Stack Rearrangement

ResearchDGX agent

arXiv:2605.17815v1 Announce Type: cross Abstract: Efficient object manipulation strategies have significant impact in automation applications. In this work, the stack rearrangement in tabletop setting

VISAFF: Speaker-Centered Visual Affective Feature Learning for Emotion Recognition in Conversation

ApplicationsDGX agent

arXiv:2605.18547v1 Announce Type: new Abstract: Emotion Recognition in Conversation (ERC) is essential for effective human-machine interaction, aiming to identify speakers' emotional states in multi-t

Vision Inference Former: Sustaining Visual Consistency in Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2605.18160v1 Announce Type: cross Abstract: In recent years, multimodal large language models (MLLMs) have achieved remarkable progress, primarily attributed to effective paradigms for integrati

Vision-OPD: Learning to See Fine Details for Multimodal LLMs via On-Policy Self-Distillation

Local AiDGX agent

arXiv:2605.18740v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) still struggle with fine-grained visual understanding, where answers often depend on small but decisive evide

Vision Transformer-Conditioned UNet for Domain-Adaptive Semantic Segmentation

SafetyDGX agent

arXiv:2605.16393v1 Announce Type: cross Abstract: Semantic segmentation is essential for analysing anatomical features in biomedical research, yet a performance gap remains for Vision Transformers (Vi

Visual Agentic Memory: Enabling Online Long Video Understanding via Online Indexing, Hierarchical Memory, and Agentic Retrieval

Model ReleasesDGX agent

arXiv:2605.16481v1 Announce Type: cross Abstract: Long video understanding requires more than large context windows. It also needs a memory mechanism that decides what visual evidence to retain, keeps

Visual Sculpting: Visually-Aligned Planning Representations for Long-Horizon Robot Clay Sculpting

SafetyDGX agent

arXiv:2605.17556v1 Announce Type: cross Abstract: Clay sculpting is a nuanced, artistic task involving dexterous manipulation with long-horizon planning to achieve high-level goals. As a robotics prob

Visual Timelines of Police Encounters in Body-Worn Camera Footage: Operational Context and Activity Cataloging for Training and Analysis in OpenBWC

ResearchDGX agent

arXiv:2605.17095v1 Announce Type: cross Abstract: Law enforcement agencies are accumulating vast amounts of body-worn camera (BWC) footage. However, this remains operationally opaque. That is, analyst

Visualizing the Invisible: Generative Visual Grounding Empowers Universal EEG Understanding in MLLMs

Model ReleasesDGX agent

arXiv:2605.18172v1 Announce Type: new Abstract: Leveraging the universal representations of pre-trained LLMs and MLLMs offers a promising path toward brain foundation models. However, visually-evoked

VLM-AutoDrive: Post-Training Vision-Language Models for Safety-Critical Autonomous Driving Events

SafetyDGX agent

arXiv:2603.18178v2 Announce Type: replace-cross Abstract: The rapid growth of ego-centric dashcam footage presents a major challenge for detecting safety-critical events such as collisions and near-co

Voice ''Cloning'' is Style Transfer

ResearchDGX agent

arXiv:2605.16578v1 Announce Type: cross Abstract: Artificially generated speech is increasingly embedded in everyday life. Voice cloning in particular enables applications where identity preservation

Voices in the Loop: Mapping Participatory AI

SafetyDGX agent

arXiv:2605.16827v1 Announce Type: new Abstract: Participatory approaches to artificial intelligence are increasingly documented across public, civic, and humanitarian settings, but evidence about how

VolTA-3D: Self-Supervised Learning for Brain MRI using 3D Volumetric Token Alignment

SafetyDGX agent

arXiv:2605.16775v1 Announce Type: cross Abstract: Self-supervised learning (SSL) has advanced medical image analysis be enabling learning form large unlabelled data. However, in brain magnetic resonan

WASIL: In-the-Wild Arabic Spoken Interactions with LLMs

ResearchDGX agent

arXiv:2605.16364v1 Announce Type: cross Abstract: Large Language Models (LLMs) voice assistants are commonly built as cascaded Automatic Speech recognition (ASR) to LLM systems, where recognition erro

Wasserstein Equilibrium Decoding for Reliable Medical Visual Question Answering

Model ReleasesDGX agent

arXiv:2605.18313v1 Announce Type: cross Abstract: Small vision-language models (2-8B) are well-suited for clin- ical deployment due to privacy constraints, limited connectivity, and low-latency requir

Watching, Reasoning, and Searching: A Video Deep Research Benchmark on Open Web for Agentic Video Reasoning

Model ReleasesDGX agent

arXiv:2601.06943v2 Announce Type: replace-cross Abstract: In real-world video question answering scenarios, videos often provide only localized visual cues, while verifiable answers are distributed ac

Wavelet Flow Matching for Multi-Scale Physics Emulation

ResearchDGX agent

arXiv:2605.16573v1 Announce Type: cross Abstract: Accurate emulation of multi-scale physical systems governed by PDEs demands models that remain stable over long autoregressive rollouts while preservi

Weak-to-Strong Elicitation via Mismatched Wrong Drafts

SafetyDGX agent

arXiv:2605.17314v1 Announce Type: cross Abstract: We consider whether off-policy experience from a smaller, weaker model can elicit capability in a stronger learner that on-policy RL fine-tuning (e.g.

WebGameBench: Requirement-to-Application Evaluation for Coding Agents via Browser-Native Games

Model ReleasesDGX agent

arXiv:2605.17637v1 Announce Type: new Abstract: Coding agents are increasingly used as application builders, yet many evaluations still focus on source code, repository-level tests, or intermediate tr

WELD: The First Naturalistic Long-Period Small-Team Workplace Emotion Dataset for Ubiquitous Affective Computing

Model ReleasesDGX agent

arXiv:2510.15221v2 Announce Type: replace Abstract: Affective computing has matured rapidly in laboratory settings, yet no prior dataset combines (i) months-to-years of duration, (ii) a naturalistic w

What Does the AI Doctor Value? Auditing Pluralism in the Clinical Ethics of Language Models

Model ReleasesDGX agent

arXiv:2605.18738v1 Announce Type: new Abstract: Medicine is inherently pluralistic. Principles such as autonomy, beneficence, nonmaleficence, and justice routinely conflict, and such ethical dilemmas

What Drives Success in Physical Planning with Joint-Embedding Predictive World Models?

ApplicationsDGX agent

arXiv:2512.24497v3 Announce Type: replace Abstract: A long-standing challenge in AI is to develop agents capable of solving a wide range of physical tasks and generalizing to new, unseen tasks and env

What is Holding Back Latent Visual Reasoning?

ResearchDGX agent

arXiv:2605.18445v1 Announce Type: cross Abstract: Humans can approach complex visual problems by mentally simulating intermediate visual steps, rather than reasoning through language alone. Inspired b

When Actions Disappear: Adversarial Action Removal in Self-Play Reinforcement Learning

AgentsDGX agent

arXiv:2605.16312v1 Announce Type: cross Abstract: We study adversarial action masking in self-play reinforcement learning: an attacker selectively removes legal actions from a victim's action set. Unl

When Bits Break Recourse: Counterfactual-Faithful Quantization

ResearchDGX agent

arXiv:2605.17160v1 Announce Type: cross Abstract: Quantization can preserve predictive accuracy under low-bit deployment while silently breaking algorithmic recourse: an actionable change that flips a

When Dynamics Shift, Robust Task Inference Wins: Offline Imitation Learning with Behavior Foundation Models Revisited

SafetyDGX agent

arXiv:2605.17017v1 Announce Type: cross Abstract: Behavior Foundation Models (BFMs) enable scalable imitation learning (IL) by pretraining task-agnostic representations that can be rapidly adapted to

When Efficiency Backfires: Cascading LLMs Trigger Cascade Failure under Adversarial Attack

ResearchDGX agent

arXiv:2605.17288v1 Announce Type: cross Abstract: Large Language Model (LLM) cascade systems are designed to balance efficiency and performance by processing queries with lightweight models while sele

When Fireflies Cluster; Enhancing Automatic Clustering via Centroid-Guided Firefly Optimization

ResearchDGX agent

arXiv:2605.18460v1 Announce Type: new Abstract: This work presents a novel variant of the Firefly Algorithm (FA) for data clustering, addressing limitations of traditional methods like K-Means that st

When Is Rank-1 Steering Cheap? Geometry, Granularity, and Budgeted Search

SafetyDGX agent

arXiv:2605.16362v1 Announce Type: cross Abstract: Activation steering offers a lightweight way to control LLMs without retraining, but its effectiveness varies sharply across concepts. Prior work ofte

When Marginals Match but Structure Fails: Covariance Fidelity in Generative Models

ResearchDGX agent

arXiv:2603.17041v2 Announce Type: replace-cross Abstract: Generative models are increasingly deployed as substitutes for real data in downstream scientific workflows, yet standard evaluation criteria

When Outcome Looks Right But Discipline Fails: Trace-Based Evaluation Under Hidden Competitor State

Model ReleasesDGX agent

arXiv:2605.18580v1 Announce Type: new Abstract: Outcome-only evaluation can certify economically unsafe agents: a policy can hit a business KPI while violating deployable behavioral discipline. In hot

When Personalization Legitimizes Risks: Uncovering Safety Vulnerabilities in Personalized Dialogue Agents

Model ReleasesDGX agent

arXiv:2601.17887v2 Announce Type: replace Abstract: Long-term memory enables large language model (LLM) agents to support personalized and sustained interactions. However, most work on personalized ag

Where Pretraining writes and Alignment reads: the asymmetry of Transformer weight space

SafetyDGX agent

arXiv:2605.16600v1 Announce Type: cross Abstract: Cross-entropy pretraining and preference alignment update the same transformer weights, but leave geometrically distinct traces. We characterise this

Whispers in the Noise: Surrogate-Guided Concept Awakening via a Multi-Agent Framework

SafetyDGX agent

arXiv:2605.18150v1 Announce Type: new Abstract: Diffusion models (DMs) are widely used for text-to-image generation, but their strong generative capabilities also raise concerns about unsafe or undesi

WhiteTesseract: Reframing the Interpretation of Cultural Heritage through XR and Conversational AI

Model ReleasesDGX agent

arXiv:2605.16972v1 Announce Type: cross Abstract: Cultural heritage exhibitions often struggle to sustain attention and support reflective engagement. Physical exhibitions rely on fixed interpretive a

Who Generated This 3D Asset? Learning Source Attribution for Generative 3D Models

Model ReleasesDGX agent

arXiv:2605.18132v1 Announce Type: cross Abstract: Generative 3D models are deployed in gaming, robotics, and immersive creation, making source attribution critical: given a 3D asset, can we identify w

Why Do Safety Guardrails Degrade Across Languages?

SafetyDGX agent

arXiv:2605.17173v1 Announce Type: cross Abstract: Large language models exhibit safety degradation in non-English languages. Standard evaluation relies on Jailbreak Success Rate (JSR), which confounds

Why Modeling Human Haptic Material Perception with AI Is Difficult

ResearchDGX agent

arXiv:2605.16602v1 Announce Type: cross Abstract: Touch plays a central role in how humans perceive and recognize materials through physical contact. Despite decades of research, the mechanisms by whi

Why We Look Where We Look: Emergent Human-like Fixations of a Foveated Visual Language Model Maximizing Scene Understanding

AgentsDGX agent

arXiv:2605.17823v1 Announce Type: cross Abstract: When humans view scenes without a specific task (free-viewing), they initially direct their eye movements toward the scene center and then fixate on p

18 May 2026

A Cascaded Generative Approach for e-Commerce Recommendations

ApplicationsDGX agent

arXiv:2605.11118v2 Announce Type: replace Abstract: Personalized storefronts in large e-commerce marketplaces are often assembled from many independent components: static themes per page section ('pla

A Few GPUs, A Whole Lotta Scale: Faithful LLM Training Emulation with PrismLLM

HardwareDGX agent

arXiv:2605.15617v1 Announce Type: cross Abstract: Large language model (LLM) training today runs on clusters spanning thousands of GPUs. While this scale enables rapid model advances, developing, debu

A Generative AI Framework for Intelligent Utility Billing CO 2 Analytics and Sustainable Resource Optimisation

SafetyDGX agent

arXiv:2605.16250v1 Announce Type: cross Abstract: Distribution utilities are now expected to deliver bills that customers can actually read attach a defensible carbon number to every kWh sold and sche

A Model Can Help Itself: Reward-Free Self-Training for LLM Reasoning

ResearchDGX agent

arXiv:2510.18814v3 Announce Type: replace-cross Abstract: Can language models improve their reasoning performance without external rewards, using only their own sampled responses for training? We show

A Scalable Multi-LLM Collaboration System with Retrieval-based Selection and Exploration-Exploitation-Driven Enhancement

Model ReleasesDGX agent

arXiv:2507.14200v2 Announce Type: replace-cross Abstract: Existing multi-LLM collaboration systems often encounter scalability challenges when integrating new LLMs and tasks, leading to suboptimal per

A Topology-Aware Spatiotemporal Handover Framework for Continuous Multi-UAV Tracking

ResearchDGX agent

arXiv:2605.15779v1 Announce Type: cross Abstract: The integration of Unmanned Aerial Vehicles(UAVs) into Intelligent Transportation Systems (ITS) offers synoptic visibility for traffic monitoring, yet

A Unified Generative-AI Framework for Smart Energy Infrastructure: Intelligent Gas Distribution, Utility Billing, Carbon Analytics, and Quantum-Inspired Optimisation

ResearchDGX agent

arXiv:2605.16232v1 Announce Type: cross Abstract: The accelerating convergence of smart metering, generative artificial intelligence, and quantum-inspired combinatorial optimisation is reshaping how e

A Unified View of Score-Based and Drifting Models

TutorialsDGX agent

arXiv:2603.07514v3 Announce Type: replace-cross Abstract: Drifting models train one-step generators by optimizing a kernel-induced mean-shift discrepancy between the data and model distributions, with

A3D: Agentic AI flow for autonomous Accelerator Design

Model ReleasesDGX agent

arXiv:2605.15237v1 Announce Type: cross Abstract: Accelerating applications through the design of hardware accelerators can significantly enhance system performance and energy efficiency. Despite adva

Access Timing as Scaffolding: A Reinforcement Learning Approach to GenAI in Education

AgentsDGX agent

arXiv:2605.15850v1 Announce Type: cross Abstract: In recent years, generative AI (GenAI) in educational settings has become ubiquitous in students' daily lives, despite its potential to induce over-re

← Previous
1…244245246247248…358
Next →