AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent
84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
2 Jun 2026

Needles at Scale: LLM-Assisted Target Selection for Windows Vulnerability Research

AgentsDGX agent

arXiv:2606.01364v1 Announce Type: cross Abstract: The attack surface of a modern operating system is a haystack: thousands of signed binaries and millions of functions, almost none relevant to any giv

Network Distributed Multi-Agent Reinforcement Learning for Consensus Control of Quadcopters

SafetyDGX agent

arXiv:2606.02107v1 Announce Type: cross Abstract: This paper proposes a Network Distributed Multi-Agent Reinforcement Learning (ND-MARL) framework for quadcopter consensus control. Compared to convent

Neural Network Compression by Approximate Differential Equivalence

Model ReleasesDGX agent

arXiv:2606.01402v1 Announce Type: cross Abstract: Neural network compression is commonly achieved by pruning parameters based on local importance scores, e.g., magnitude-based pruning. We propose a co


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

NILC: Discovering New Intents with LLM-assisted Clustering

Model ReleasesDGX agent

arXiv:2511.05913v2 Announce Type: replace-cross Abstract: New intent discovery (NID) seeks to recognize both new and known intents from unlabeled user utterances, which finds prevalent use in practica

Not All Errors Are Equal: A Systematic Study of Error Propagation in Large Language Model Inference

ResearchDGX agent

arXiv:2606.02430v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly integrated into high-performance computing (HPC) workflows, accelerating scientific discovery through di

OctoT2I: A Self-Evolving Agentic Text-to-Image Router

AgentsDGX agent

arXiv:2606.01803v1 Announce Type: new Abstract: The explosive growth of Text-to-Image (T2I) models, from large-scale versions to lightweight, real-time ones, now faces diminishing marginal returns fro

ODTQA-FoRe: An Open-Domain Tabular Question Answering Dataset for Future Data Forecasting and Reasoning

AgentsDGX agent

arXiv:2606.02433v1 Announce Type: cross Abstract: The rapid development of LLMs has significantly advanced tabular question answering, but most systems cannot perform future-oriented numerical predict

Off-the-Shelf LLMs as Process Scorers: Training-Free Alternative to PRMs for Mathematical Reasoning

Model ReleasesDGX agent

arXiv:2606.01682v1 Announce Type: cross Abstract: Selecting the best response from multiple small-model samples using a stronger scorer is a simple inference-time strategy, but fails when the small mo

On Effectiveness and Efficiency of Agentic Tool-calling and RL Training

SafetyDGX agent

arXiv:2606.00135v1 Announce Type: cross Abstract: Tool-calling is a central component of modern large language model (LLM) agents, equipping them with skills beyond their parametric knowledge. This pa

On Imbalanced Regression with Hoeffding Trees

ApplicationsDGX agent

arXiv:2602.22101v3 Announce Type: replace-cross Abstract: Many real-world applications generate continuous data streams for regression. Hoeffding trees and their variants have a long-standing traditio

On Information Self-Locking in Reinforcement Learning for Active Reasoning of LLM agents

AgentsDGX agent

arXiv:2603.12109v2 Announce Type: replace Abstract: Reinforcement learning (RL) has become a de facto paradigm for building LLM-based agents that act, interact, and reason over extended task horizons.

On the Collapse of Generative Paths: A Criterion and Correction for Diffusion Steering

ResearchDGX agent

arXiv:2512.10339v2 Announce Type: replace Abstract: Inference-time steering adapts pretrained diffusion and flow models to new tasks without retraining, often utilizing ratio-of-densities construction

On the Difficulty of Learning a Meta-network for Training Data Selection

TutorialsDGX agent

arXiv:2606.00571v1 Announce Type: cross Abstract: Synthetic data are increasingly used to train neural networks, yet distributional mismatch with real data limits their effectiveness when used indiscr

On the Evaluation of Spiking Neural Network Configurations for Network Intrusion Detection

Model ReleasesDGX agent

arXiv:2606.01442v1 Announce Type: cross Abstract: Network intrusion detection is a core component of modern cybersecurity infrastructure, yet the deep learning models that dominate the field are compu

On the evolution of the concept of probability as a mirror of the evolution of reason

ResearchDGX agent

arXiv:2606.00102v1 Announce Type: new Abstract: Over the centuries, probability theory has grown from the calculus of games of chance into a central framework for reasoning under uncertainty. This art

On the Generalization in Topology Optimization via Sensitivity-Conditioned Bernoulli Flow Matching

Model ReleasesDGX agent

arXiv:2606.02179v1 Announce Type: cross Abstract: Surrogate models for topology optimization (TO) exhibit highly variable out-of-distribution (OOD) generalization under distribution shifts such as cha

On the Limits of LLM Adaptability: Impact of Model-Internalized Priors on Annotation Task Performance

SafetyDGX agent

arXiv:2606.00467v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used for zero-shot annotation and LLM-as-a-judge tasks, yet their reliability hinges on how model-intern

On the Limits of Token Reduction for Efficient Unified Vision Language Training

Model ReleasesDGX agent

arXiv:2606.01503v1 Announce Type: cross Abstract: Unified vision-language models (VLMs) integrate visual understanding and visual generation within a single autoregressive backbone, but their joint tr

On the Theoretical Limitations of Embedding-based Link Prediction

Model ReleasesDGX agent

arXiv:2506.22271v3 Announce Type: replace Abstract: Neural networks often map low-dimensional embeddings to high-dimensional output spaces. Usually, the output layer is linear, which can create a 'ran

On Wednesdays, We Ask Questions: Optimizing 'Active Listening' in Automated Legal Triage and Referral

Model ReleasesDGX agent

arXiv:2606.00272v1 Announce Type: new Abstract: The FETCH classifier generates follow-up questions to help refine the best match for the applicant's legal problem, using a low-cost ensemble of LLMs. I

One Bias After Another: Mechanistic Reward Shaping and Persistent Biases in Language Reward Models

SafetyDGX agent

arXiv:2603.03291v2 Announce Type: replace-cross Abstract: Reward Models (RMs) are crucial for online alignment of language models (LMs) with human preferences. However, RM-based preference-tuning is v

OPD+: Rethinking the Advantage Design for On-Policy Distillation

SafetyDGX agent

arXiv:2606.01039v1 Announce Type: cross Abstract: On-policy distillation (OPD) is a widely used technique to transfer capabilities from capable teacher language models to the base student models, and

OpenHospital: A Thing-in-itself Arena for Evolving and Benchmarking LLM-based Collective Intelligence

AgentsDGX agent

arXiv:2603.14771v3 Announce Type: replace Abstract: Large Language Model (LLM)-based Collective Intelligence (CI) presents a promising approach to overcoming the data wall and continuously boosting th

OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents

Model ReleasesDGX agent

arXiv:2606.02031v1 Announce Type: cross Abstract: Building capable visual web agents requires long-horizon reasoning, precise grounding, and robust interaction with dynamic real-world websites. Despit

Optimal Bayesian Stopping for Efficient Inference of Consistent LLM Answers

SafetyDGX agent

arXiv:2602.05395v2 Announce Type: replace-cross Abstract: A simple strategy for improving LLM accuracy, especially in math and reasoning problems, is to sample multiple responses and submit the answer

Optimal Transport-based Permutation-Invariant Bayesian Optimization of Offshore Wind Farm Layouts

ApplicationsDGX agent

arXiv:2606.00009v1 Announce Type: new Abstract: Bayesian Optimization (BO) is widely and successfully adopted for solving optimization problems having an expensive-to-evaluate, black-box, and non-conv

Optimizing Diversity and Quality through Base-Aligned Model Collaboration

SafetyDGX agent

arXiv:2511.05650v2 Announce Type: replace-cross Abstract: Alignment has greatly improved large language models (LLMs)' output quality at the cost of diversity, yielding highly similar outputs across g

Order within Chaos: Capturing Intrinsic Energy Anomalies for AI-Manipulated Image Forgery Localization

Model ReleasesDGX agent

arXiv:2606.02178v1 Announce Type: cross Abstract: Recent advancements in generative AI have led to image editing models capable of producing realistic forgeries that evade traditional image forgery lo

PaCo-VLA: Passivity-Shielded Compliance Prior for Contact-Rich Vision-Language-Action Manipulation

ApplicationsDGX agent

arXiv:2606.00515v1 Announce Type: cross Abstract: Contact-rich manipulation demands both high-level semantic reasoning and the safe regulation of high-frequency contact dynamics. While Vision-Language

PALTO: Physics-Informed Active Learning for Tri-Gate FinFET Design Optimization for Vertical Power Delivery

TutorialsDGX agent

arXiv:2606.01265v1 Announce Type: cross Abstract: This paper demonstrates the effectiveness of machine learning-driven optimization for designing application-specific GaN tri-gate FinFETs in vertical

Paradoxical noise preference in RNNs

SafetyDGX agent

arXiv:2601.04539v2 Announce Type: replace-cross Abstract: In recurrent neural networks (RNNs) used to model biological neural networks, noise is typically introduced during training to emulate biologi

Parameter-Efficient Fine-Tuning of Large Pretrained Models for Instance Segmentation Tasks

Model ReleasesDGX agent

arXiv:2606.01947v1 Announce Type: cross Abstract: Research and applications in artificial intelligence have recently shifted with the rise of large pretrained models, which deliver state-of-the-art re

PaSBench-Video: A Streaming Video Benchmark for Proactive Safety Warning

Model ReleasesDGX agent

arXiv:2606.02443v1 Announce Type: cross Abstract: Between the first visible sign of danger and the moment an accident occurs, there is often a window where intervention remains possible. Video-capable

Pause and Think: A Dataset and Benchmark for Video-Grounded Assistive Action Suggestion

Model ReleasesDGX agent

arXiv:2606.00616v1 Announce Type: cross Abstract: Recent Vision-Language Models (VLMs) struggle with grounded reasoning, temporal consistency, and context aware planning in videos. We introduce pause-

pcbGPT: Automatic PCB Schematic Synthesis from Natural Language Requirements

ResearchDGX agent

arXiv:2606.01188v1 Announce Type: cross Abstract: Translating natural-language hardware requirements into correct printed circuit board (PCB) schematics remains difficult in embedded, IoT, and wearabl

PEACE: A Planner-Executor Agent with Constraint Enforcement for UAVs

Local AiDGX agent

arXiv:2606.00104v1 Announce Type: cross Abstract: Foundation models are increasingly used to drive autonomous systems, yet existing approaches either keep the model in a tight control loop, raising la

PECKER: A Precisely Efficient Critical Knowledge Erasure Recipe For Machine Unlearning in Diffusion Models

ResearchDGX agent

arXiv:2604.05634v2 Announce Type: replace Abstract: Machine unlearning (MU) has become a critical technique for GenAI models' safe and compliant operation. While existing MU methods are effective, mos

Permissive Safety Through Trusted Inference: Verifiable Belief-Space Neural Safety Filters for Assured Interactive Robotics

Model ReleasesDGX agent

arXiv:2606.02562v1 Announce Type: cross Abstract: Autonomous robots that interact with people must make safe and efficient decisions under human-induced uncertainty, such as their preferences, goals,

Persona Attack: Incremental Memory Injection Jailbreak Attack against Large Language Models

Model ReleasesDGX agent

arXiv:2606.00150v1 Announce Type: cross Abstract: As Large Language Models evolve for user convenience, vulnerability to jailbreak attacks continues to be reported despite ongoing efforts in safety tr

Perspective on Bias in Biomedical AI: Preventing Downstream Healthcare Disparities

SafetyDGX agent

arXiv:2604.14514v2 Announce Type: replace Abstract: Healthcare disparities persist across socioeconomic boundaries, often attributed to unequal access to screening, diagnostics, and therapeutics. Howe

Perturbation Effects on Accuracy and Fairness among Similar Individuals

SafetyDGX agent

arXiv:2404.01356v3 Announce Type: replace-cross Abstract: Deep neural networks are vulnerable to adversarial perturbations that can simultaneously degrade prediction robustness and individual fairness

PETS: A Principled Framework Towards Optimal Trajectory Allocation for Efficient Test-Time Self-Consistency

ResearchDGX agent

arXiv:2602.16745v2 Announce Type: replace-cross Abstract: Test-time scaling can improve model performance by aggregating stochastic reasoning trajectories. However, achieving sample-efficient test-tim

Physically-Constrained Mamba-SDE for Remaining Useful Life Prediction under Irregular Observations

SafetyDGX agent

arXiv:2606.01894v1 Announce Type: new Abstract: Accurate Remaining Useful Life prediction is critical for industrial predictive maintenance. However, real-world deployment is challenging due to the ir

Physics-Encoded Inverse Modeling for Arctic Snow Depth Prediction

Model ReleasesDGX agent

arXiv:2601.17074v4 Announce Type: replace-cross Abstract: Accurate estimation in time-varying inverse problems under limited and sparse observations remains a fundamental challenge across scientific d

Physics-Guided Attention in a Lightweight TCN for Efficient WiFi CSI-Based Human Activity Recognition

Model ReleasesDGX agent

arXiv:2606.01834v1 Announce Type: cross Abstract: Human Action Recognition (HAR) using WiFi Channel State Information (CSI) has gained increasing attention due to its non-contact, low-cost, and privac

Physics-Informed Deep Learning for Entropy Prediction in Heterogeneous Systems: Thermodynamic and Information-Theoretic Case Studies

ApplicationsDGX agent

arXiv:2606.01179v1 Announce Type: cross Abstract: Entropy production governs irreversibility and uncertainty in both physical and information-theoretic systems. While Physics-Informed Neural Networks

Physics-Informed Neural Networks for Radial Consolidation of Combined Electroosmotic, Vacuum and Surcharge Preloading Considering Smear Effects

TutorialsDGX agent

arXiv:2606.00056v1 Announce Type: cross Abstract: This study develops a dimensionless multi-domain physics-informed neural network (PINN) framework for electro-osmotic radial consolidation considering

PlanarBench: Evaluating LLM Spatial Reasoning via Planar Graph Drawing

ResearchDGX agent

arXiv:2606.02010v1 Announce Type: cross Abstract: PlanarBench tests whether LLMs can draw planar graphs as ASCII art given only an edge list -- a spatial reasoning task that resists memorization becau

Planktonzilla: Multimodal dataset and models for understanding plankton ecosystems

ResearchDGX agent

arXiv:2606.00080v1 Announce Type: cross Abstract: Marine plankton underpin aquatic food webs and play a key role in global CO2 sequestration, making reliable species identification critical for unders

Plausibility Is Not Prediction: Contrastive Evidence for LLM-Based Cellular Perturbation Reasoning

ResearchDGX agent

arXiv:2606.01042v1 Announce Type: cross Abstract: Perturbation experiments are central to understanding cellular mechanisms, but remain costly and sparse, motivating prediction of gene expression resp

POIROT: Interrogating Agents for Failure Detection in Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2606.02282v1 Announce Type: new Abstract: Orchestrating Large Language Models into Multi-Agent Systems (LLM-MAS) has unlocked remarkable reasoning capabilities, yet emergent failures and halluci

PolarMem: A Training-Free Polarized Latent Graph Memory for Verifiable Vision-Language Models

ResearchDGX agent

arXiv:2602.00415v2 Announce Type: replace Abstract: Memory is not merely a storage mechanism for intelligent systems, but a structure for organizing evidence and constraining belief. This is especiall

Policy and World Modeling Co-Training for Language Agents

SafetyDGX agent

arXiv:2606.02388v1 Announce Type: cross Abstract: Reinforcement learning (RL) improves large language model (LLM) agents by teaching them which actions lead to high rewards, but provides little superv

PolySpeech-100: A Large-Scale Benchmark for Speech Understanding Across 100+ Languages and Dialects

Model ReleasesDGX agent

arXiv:2606.01016v1 Announce Type: cross Abstract: While End-to-End (E2E) Speech-Large Language Models (Speech-LLMs) are rapidly evolving, their evaluation methodologies remain limited to the era of si

Position: Beyond Sensitive Attributes, ML Fairness Should Quantify Structural Injustice via Social Determinants

SafetyDGX agent

arXiv:2508.08337v3 Announce Type: replace-cross Abstract: Algorithmic fairness research has largely framed unfairness as discrimination along sensitive attributes. However, this approach limits visibi

Position Paper: Post-Solve Robustness in Decision Engines: Feasible Regions and Smoothness Under Perturbations

Model ReleasesDGX agent

arXiv:2606.00002v1 Announce Type: new Abstract: Mixed-Integer Linear Programming (MILP) decision engines routinely output nominally optimal plans for high-stakes industrial systems. Yet deployment rar

Post-Deterministic Distributed Systems: A New Foundation for Trustworthy Autonomous Infrastructure

SafetyDGX agent

arXiv:2606.01722v1 Announce Type: cross Abstract: For decades, distributed systems have typically assumed that correct participants execute protocol-specified behavior with stable, externally defined,

PR2: Predictive Routing Replay for MoE-Based LLM Reinforcement Learning

SafetyDGX agent

arXiv:2606.00395v1 Announce Type: cross Abstract: Mixture of Experts (MoE) Large Language Models (LLMs) achieve strong performance at scale. However, reinforcement learning (RL) on MoE-based LLMs ofte

Pre-Deployment Robustness Stress Testing for CT Segmentation Systems Using Clinically Motivated Multi-Corruption Augmentation

Model ReleasesDGX agent

arXiv:2606.00491v1 Announce Type: cross Abstract: Deep learning-based CT segmentation systems often achieve high accuracy on clean benchmark images, but their performance may degrade under heterogeneo

Predicting Future Utility: Global Combinatorial Optimization for Task-Agnostic KV Cache Eviction

HardwareDGX agent

arXiv:2602.08585v2 Announce Type: replace-cross Abstract: Given the quadratic complexity of attention, KV cache eviction is vital to accelerate model inference. Current KV cache eviction methods typic

← Previous
1…177178179180181…358
Next →