AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
25 May 2026

Classical State Preparation for Variational Quantum Algorithms via Reinforcement Learning

SafetyDGX agent

arXiv:2605.23138v1 Announce Type: cross Abstract: Variational Quantum Algorithms (VQAs) potentially offer a pathway to practical quantum advantage, but their optimization is heavily hindered by barren

Co-ReAct: Rubrics as Step-Level Collaborators for ReAct Agents

AgentsDGX agent

arXiv:2605.23590v1 Announce Type: new Abstract: ReAct-style agents for search-intensive, multi-step reasoning tasks rely largely on their own internal judgment to decide what evidence to seek, which r

Coloring the Noise: Adversarial Sobolev Alignment for Faithful Image Super Resolution

SafetyDGX agent

arXiv:2605.23264v1 Announce Type: cross Abstract: Generative priors in Image Super-Resolution (SR) often compromise faithful restoration, we attribute this limitation to a fundamental spectral misalig


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Computable Fairness: Boltzmann-Softmax Control for AI Resource Allocation

Model ReleasesDGX agent

arXiv:2605.22827v1 Announce Type: cross Abstract: In large-scale AI systems, allocating scarce resources such as GPU compute time and bandwidth among multiple agents is a critical challenge. Conventio

ConjNorm: Tractable Density Estimation for Out-of-Distribution Detection

ResearchDGX agent

arXiv:2402.17888v5 Announce Type: replace-cross Abstract: Post-hoc out-of-distribution (OOD) detection has garnered intensive attention in reliable machine learning. Many efforts have been dedicated t

Controlled Personalization in Legacy Media Online Services: A Case Study in News Recommendation

SafetyDGX agent

arXiv:2510.09136v2 Announce Type: replace-cross Abstract: Personalized news recommendations have become a standard feature of large news aggregation services, optimizing user engagement through automa

Convergence Without Understanding: When Language Models Agree on Representations but Disagree on Reasoning

SafetyDGX agent

arXiv:2605.23315v1 Announce Type: cross Abstract: Large language models trained under diverse objectives and architectures have been shown to develop increasingly similar internal representations, an

CoReVAD: A Contextual Reasoning Framework for Training-Free Video Anomaly Detection

SafetyDGX agent

arXiv:2605.23116v1 Announce Type: cross Abstract: Existing Video Anomaly Detection (VAD) methods typically rely on task-specific training, leading to strong domain dependency and high training costs.

CoSPlay: Cooperative Self-Play at Test-Time with Self-Generated Code and Unit Test

ResearchDGX agent

arXiv:2605.23491v1 Announce Type: cross Abstract: Recently, Reinforcement Learning with Verifiable Rewards (RLVR) and Test-Time Scaling (TTS) have advanced LLM code generation through executable verif

Cost-Effective Model Evaluation with Meta-Learning

Model ReleasesDGX agent

arXiv:2605.23595v1 Announce Type: cross Abstract: The rapid growth of machine learning has produced an ever-expanding ecosystem of models, making it increasingly challenging to verify the reliability

CP or DP? Why Not Both: A Case Study in the Partial Shop Scheduling Problem

ApplicationsDGX agent

arXiv:2605.23569v1 Announce Type: new Abstract: Dynamic Programming (DP) and Constraint Programming (CP) are well-established paradigms for solving combinatorial optimization problems. Usually, these

Curriculum reinforcement learning with measurable task representation learning

AgentsDGX agent

arXiv:2605.23372v1 Announce Type: cross Abstract: In curriculum reinforcement learning (CRL), an agent incrementally accumulates knowledge over a sequence of tasks (i.e., a curriculum), and the learni

CVSearch: Empowering Multimodal LLMs with Cognitive Visual Search for High-Resolution Image Perception

Model ReleasesDGX agent

arXiv:2605.23655v1 Announce Type: cross Abstract: High-resolution (HR) image perception presents a key bottleneck for multimodal large language models (MLLMs). While visual search offers a promising s

DART: Semantic Recoverability for Structured Tool Agents

Local AiDGX agent

arXiv:2605.23311v1 Announce Type: new Abstract: When a structured tool agent fails mid-execution, the runtime faces a dilemma: replaying the entire task is safe but wasteful, while restoring from a lo

Decomposing and Measuring Evaluation Awareness

Model ReleasesDGX agent

arXiv:2605.23055v1 Announce Type: cross Abstract: Frontier language models sometimes recognize that they are being evaluated and adjust their behavior, undermining validity of benchmark results. Yet t

Defining AI Fatigue in Academic Contexts: Dimensions, Indicators, and a Stage-Based Model Using Grounded Theory

ResearchDGX agent

arXiv:2605.23123v1 Announce Type: cross Abstract: The integration of AI tools in academic settings has introduced a distinct form of strain that existing frameworks like technostress and digital fatig

Deja Vu in Plots: Leveraging Cross-Session Evidence with Retrieval-Augmented LLMs for Live Streaming Risk Assessment

Local AiDGX agent

arXiv:2601.16027v2 Announce Type: replace Abstract: The rise of live streaming has transformed online interaction, enabling massive real-time engagement but also exposing platforms to complex risks su

Design and Report Benchmarks for Knowledge Work

Model ReleasesDGX agent

arXiv:2605.23262v1 Announce Type: new Abstract: The development of LLM agents has led to a growing body of work on knowledge-work AI, including coding, research, and healthcare. However, current knowl

Diffusion and Flow Matching Models for Tabular Data: A Survey

SafetyDGX agent

arXiv:2502.17119v2 Announce Type: replace-cross Abstract: Deep generative models have made rapid progress in image, text, audio, and video generation, and are increasingly being applied to structured

DiLaDiff: Distilled Latent-Augmented Diffusion for Language Modeling

ResearchDGX agent

arXiv:2605.23605v1 Announce Type: cross Abstract: Diffusion language models intrinsically fail to capture correlations between decoded tokens, which leads to a harsh trade-off between sampling quality

Disentangling Interaction and Bias Effects in Opinion Dynamics of Large Language Models

SafetyDGX agent

arXiv:2509.06858v2 Announce Type: replace-cross Abstract: Large Language Models are increasingly used to simulate human opinion dynamics, yet the effect of genuine interaction is often obscured by sys

Dithering Defense: Adversarial Robustness of Vision Foundation Models via Multi-Level Floyd-Steinberg Dithering

ResearchDGX agent

arXiv:2605.23065v1 Announce Type: cross Abstract: Vision foundation models are widely used as frozen backbones across many downstream tasks, making them a single point of failure under adversarial att

Do Language Models Know What Not to Say? Causal Evidence for Statistical Preemption in LLMs

ResearchDGX agent

arXiv:2605.23039v1 Announce Type: cross Abstract: How do learners acquire knowledge of what is unacceptable without negative evidence? Construction Grammar proposes statistical preemption: exposure to

Do Synthetic Brain MRIs Reliably Improve Tumour Classification? A StyleGAN2-ADA Class-Plane Augmentation Study on BRISC 2025

Model ReleasesDGX agent

arXiv:2605.23094v1 Announce Type: cross Abstract: Generative augmentation is often proposed as a remedy for small medical-image datasets, but synthetic images are only useful when they improve downstr

DrawVideo: Generating Long Video from Storyboard Keyframe Sketches

ResearchDGX agent

arXiv:2605.23508v1 Announce Type: cross Abstract: Long video generation requires high-fidelity synthesis, coherent narrative structure, and user control over extended time spans. Existing text-to-vide

DreamerNLplus: Interpretable Modeling of Mental Health Dynamics from Social Media Timelines using Hybrid Rule-Based and RAG Methods

Model ReleasesDGX agent

arXiv:2605.23052v1 Announce Type: cross Abstract: We present DreamerNLplus, a hybrid framework for modeling mental health dynamics from social media timelines in the CLPsych 2026 shared task. Our syst

Dreaming Smoothly and Sample Efficiently with Gradient Penalized Latent Dynamics

SafetyDGX agent

arXiv:2605.23089v1 Announce Type: cross Abstract: Model-based reinforcement learning improves sample efficiency by learning a world model. However, existing latent world models such as DreamerV3 do no

DRL-Driven Edge-Aware Utility Optimization for Multi-Slice 6G Networks

ResearchDGX agent

arXiv:2605.23056v1 Announce Type: cross Abstract: Virtual Reality (VR) services delivered over 6G networks demand ultra-low latency and high bandwidth to ensure seamless user experiences. This paper p

DualMem: Bypassing the Objectness Bottleneck for Calibrated Unknown-Stream Filtering in Open-World Object Detection

ResearchDGX agent

arXiv:2605.23634v1 Announce Type: cross Abstract: Open-world object detection (OWOD) requires detectors to localize known classes while identifying unknown objects for future incremental learning. We

EDGE-OPD: Internalizing Privileged Context with Evidence Guided On-Policy Distillation

SafetyDGX agent

arXiv:2605.23493v1 Announce Type: new Abstract: On-Policy Distillation (OPD) has gained wide attraction as an LLM post-training paradigm due to its effectiveness in improving capabilities without intr

Efficient and Transferable Agentic Knowledge Graph RAG via Reinforcement Learning

Model ReleasesDGX agent

arXiv:2509.26383v5 Announce Type: replace-cross Abstract: Knowledge-graph retrieval-augmented generation (KG-RAG) couples large language models (LLMs) with structured, verifiable knowledge graphs (KGs

EM-Vid: Training-Free Entity-Centric Memory for Efficient and Consistent Multi-Shot Video Generation

ResearchDGX agent

arXiv:2605.23610v1 Announce Type: cross Abstract: Multi-shot video generation requires maintaining a consistent appearance of recurring entities across shots while remaining faithful to shot-specific

Empowering 9-1-1 Calltaking Training with Generative AI: Experiences and Lessons Learned

SafetyDGX agent

arXiv:2602.13241v2 Announce Type: replace-cross Abstract: Emergency call-takers form the first operational link in public safety response, handling over 240 million calls annually while facing a susta

Energy per Successful Goal: Goal-Level Energy Accounting for Agentic AI Systems

SafetyDGX agent

arXiv:2605.22883v1 Announce Type: new Abstract: Current AI energy benchmarks measure consumption at the granularity of a single model invocation or training run. For classical single-turn workloads th

Enhancing Deep Neural Network Reliability with Refinement and Calibration

ResearchDGX agent

arXiv:2605.23249v1 Announce Type: cross Abstract: Although deep neural networks (DNNs) achieve high predictive accuracy, their confidence estimates are often unreliable, potentially compromising user

ETCHR: Editing To Clarify and Harness Reasoning

Model ReleasesDGX agent

arXiv:2605.23897v1 Announce Type: cross Abstract: Multimodal Large Language Models have advanced visual reasoning, yet a purely textual chain of thought remains a bottleneck for questions that require

Evaluating Large Language Models in a Complex Hidden Role Game

Model ReleasesDGX agent

arXiv:2605.22826v1 Announce Type: cross Abstract: Quantifying the deceptive potential of Large Language Models (LLMs) is critical for AI safety, yet difficult to achieve in uncontrolled environments.

EvalVerse: Pipeline-Aware and Expert-Calibrated Benchmarking for Professional Cinematic Video Generation

AgentsDGX agent

arXiv:2605.23271v1 Announce Type: cross Abstract: The rapid evolution of generative video foundation models has propelled the field toward professional-grade cinematic synthesis. To achieve such deman

EVE-Agent: Evidence-Verifiable Self-Evolving Agents

AgentsDGX agent

arXiv:2605.22905v1 Announce Type: new Abstract: Self-evolving agents should not train on examples they cannot justify. Data-free self-evolving search agents offer a scalable route to systems that gene

Every Component is a Lookup: Token Attribution and Composition from a Single Decomposition

ResearchDGX agent

arXiv:2605.23393v1 Announce Type: cross Abstract: Mechanistic interpretability of transformers requires identifying not just which components matter but how they compose into the computational route t

Exploiting Longitudinal Context in Clinician-Verified Interactive Lesion Tracking

Model ReleasesDGX agent

arXiv:2605.23118v1 Announce Type: cross Abstract: Tracking tumor lesions across serial CT scans is essential for oncological response assessment. Existing automated methods face a fundamental trade-of

Expressive Power of Deep Homomorphism Networks over Relational Databases

ResearchDGX agent

arXiv:2605.22852v1 Announce Type: cross Abstract: The expressive limitations of message-passing Graph Neural Networks (GNNs) have motivated a wide range of more powerful graph learning architectures.

FastKernels: Benchmarking GPU Kernel Generation in Production

Model ReleasesDGX agent

arXiv:2605.23215v1 Announce Type: cross Abstract: LLM-based agents for GPU kernel generation are advancing rapidly, yet their progress is fundamentally constrained by the benchmarks they optimize agai

FATHOMS-RAG: A Framework for the Assessment of Thinking and Observation in Multimodal Systems that use Retrieval Augmented Generation

Model ReleasesDGX agent

arXiv:2510.08945v3 Announce Type: replace Abstract: Retrieval-augmented generation (RAG) has emerged as a promising paradigm for improving factual accuracy in large language models (LLMs). We introduc

Fine-Tuning Causal LLMs for Text Classification: Embedding-Based vs. Instruction-Based Approaches

Model ReleasesDGX agent

arXiv:2512.12677v2 Announce Type: replace-cross Abstract: We explore efficient strategies to fine-tune decoder-only Large Language Models (LLMs) for downstream text classification under resource const

Forget What's Sensitive, Remember What Matters: Token-Level Differential Privacy in Memory Sculpting for Continual Learning

ResearchDGX agent

arXiv:2509.12958v2 Announce Type: replace Abstract: Continual Learning (CL) models, while adept at sequential knowledge acquisition, face significant and often overlooked privacy challenges due to acc

Foundation Protocol: A Coordination Layer for Agentic Society

SafetyDGX agent

arXiv:2605.23218v1 Announce Type: new Abstract: Autonomous agents are moving from tools into a layer of social infrastructure: they browse, purchase, deploy software, manage systems, and increasingly

From Raw Experience to Skill Consumption: A Systematic Study of Model-Generated Agent Skills

AgentsDGX agent

arXiv:2605.23899v1 Announce Type: new Abstract: Language agents increasingly improve by reusing skills -- structured procedural artifacts distilled from past experience. In particular, domain-level an

Generative AI and the Reorganization of Labor Demand

ResearchDGX agent

arXiv:2605.23159v1 Announce Type: cross Abstract: Generative artificial intelligence (AI) is expected to transform work, but less is known about how firms reorganize labor demand as the technology dif

GENSTRAT: Toward a Science of Strategic Reasoning in Large Language Models

Model ReleasesDGX agent

arXiv:2605.23238v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed as economic agents in marketplaces, auctions, and bidding settings. Anticipating their behavior i

GeoMAE: Masking Representation Learning for Spatio-Temporal Graph Forecasting with Missing Values

ApplicationsDGX agent

arXiv:2508.14083v3 Announce Type: replace-cross Abstract: The ubiquity of missing data in urban intelligence systems, attributable to adverse environmental conditions and equipment failures, poses a s

GILT: An LLM-Free, Tuning-Free Graph Foundational Model for In-Context Learning

ResearchDGX agent

arXiv:2510.04567v2 Announce Type: replace-cross Abstract: Graph Neural Networks (GNNs) are powerful tools for processing relational data but often struggle to generalize to unseen graphs, giving rise

GlyTwin: Digital Twin for Glucose Control in Type 1 Diabetes Through Optimal Behavioral Modifications Using Patient-Centric Counterfactuals

TutorialsDGX agent

arXiv:2504.09846v2 Announce Type: replace-cross Abstract: Frequent and long-term exposure to hyperglycemia increases the risk of chronic complications, including neuropathy, nephropathy, and cardiovas

Goal-Conditioned Agents that Learn Everything All at Once

SafetyDGX agent

arXiv:2605.23551v1 Announce Type: cross Abstract: A goal-conditioned reinforcement learning agent exploring an environment will see a wealth of information throughout a trajectory, most of which is di

Good Token Hunting: A Hitchhiker's Guide to Token Selection for Visual Geometry Transformers

TutorialsDGX agent

arXiv:2605.23892v1 Announce Type: cross Abstract: Visual geometry transformers have become powerful architectures for multi-view 3D reconstruction, enabling joint prediction of multiple 3D attributes

GradingAttack: Exposing Security Vulnerabilities in LLM Based Educational Grading Agents

AgentsDGX agent

arXiv:2602.00979v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly deployed as educational agents for automatic short answer grading (ASAG) in real-world education

Graph Alignment Topology as an Inductive Bias for Grounding Detection

SafetyDGX agent

arXiv:2605.22963v1 Announce Type: cross Abstract: Large Language Models (LLMs) are optimized to produce distributionally plausible continuations rather than to explicitly verify whether generated prop

GT-HarmBench: Benchmarking AI Safety Risks Through the Lens of Game Theory

Model ReleasesDGX agent

arXiv:2602.12316v2 Announce Type: replace Abstract: Frontier AI systems are increasingly capable and deployed in high-stakes multi-agent environments. However, existing AI safety benchmarks largely ev

HARNESS-LM: A Three-Phase Training Recipe for Harnessing SLMs in Sponsored Search Retrieval

Model ReleasesDGX agent

arXiv:2605.23572v1 Announce Type: cross Abstract: In the competitive landscape of sponsored search, balancing retrieval quality with production latency is a critical challenge. While large retrieval m

How Far Will They Go? Red-Teaming Online Influence with Large Language Models

Local AiDGX agent

arXiv:2605.22880v1 Announce Type: cross Abstract: As large language model (LLM)-based agents increasingly participate in online discourse, red-teaming their capacity to support political influence cam

← Previous
1…221222223224225…358
Next →