AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent
84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
2 Jun 2026

STaR-KV: Spatio-Temporal Adaptive Re-weighting for KV Cache Compression in GUI Vision-Language Models

HardwareDGX agent

arXiv:2606.01790v1 Announce Type: cross Abstract: Vision-language-model-based graphical user interface (GUI) agents have shown broad automation capabilities, yet deployment is bottlenecked by a key-va

STARFISH: faST Accuracy Recovery in pruned networks From Internal State Healing

ResearchDGX agent

arXiv:2606.01126v1 Announce Type: cross Abstract: Pruning is a process designed to reduce the number of weights in a large neural network. This can substantially speed up inference but might cause a c

StemBind: When MLLMs Get Lost Between Rules and Instances in Abstract Visual Reasoning

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.00148v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) often know the rule but pick the wrong answer: on abstract visual reasoning (AVR) tasks, a model can describe

Stochastic convergence of parallel asynchronous adaptive first-order methods

ResearchDGX agent

arXiv:2606.01787v1 Announce Type: new Abstract: A new class of asynchronous adaptive first-order optimization methods is introduced, comprising asynchronous variants of several popular algorithms. Ver

Stop Wandering, Find the Keys: LLMs Discriminate Key States for Efficient Multi-Agent Exploration

AgentsDGX agent

arXiv:2410.02511v2 Announce Type: replace Abstract: With expansive state-action spaces, efficient multi-agent exploration remains a longstanding challenge in reinforcement learning. Although pursuing

StreamingVLM: Real-Time Understanding for Infinite Video Streams

Model ReleasesDGX agent

arXiv:2510.09608v2 Announce Type: replace-cross Abstract: Vision-language models (VLMs) could power real-time assistants and autonomous agents, but they face a critical challenge: understanding near-i

StressDream: Steering Video World Models for Robust Policy Evaluation and Improvement

SafetyDGX agent

arXiv:2606.00267v1 Announce Type: cross Abstract: Video world models (WMs) have shown promise for policy evaluation and improvement by imagining realistic future observations conditioned on ego-robot

Strong Stochastic Flow Maps

TutorialsDGX agent

arXiv:2606.01086v1 Announce Type: cross Abstract: Flow and diffusion models generate high-quality samples in many modalities; however, many network evaluations are required during inference due to num

Structure Enables Effective Self-Localization of Errors in LLMs

Local AiDGX agent

arXiv:2602.02416v2 Announce Type: replace Abstract: Self-correction in language models remains elusive. In this work, we explore whether language models can explicitly localize errors in incorrect rea

Structure-Guided Adaptive Propagation for Protein-Protein Interaction Site Prediction

Local AiDGX agent

arXiv:2606.01781v1 Announce Type: new Abstract: Accurate prediction of protein-protein interaction sites (PPIS) is essential for understanding cellular processes, disease mechanisms, and therapeutic t

Structured Visual Evidence Decomposition for Evidence-Grounded Multimodal Screening of Obstructive Sleep Apnea-Hypopnea Syndrome

ResearchDGX agent

arXiv:2606.00087v1 Announce Type: cross Abstract: Effective pre-polysomnography screening for obstructive sleep apnea-hypopnea syndrome (OSAHS) requires combining clinical risk factors with visible cr

Subliminal Learning is a LoRA Artifact

Model ReleasesDGX agent

arXiv:2606.00831v1 Announce Type: new Abstract: Subliminal learning is a phenomenon where language models can transmit behavioral traits to other models through seemingly innocuous data (Cloud et al.,

Subliminal Learning Is Steering Vector Distillation

ResearchDGX agent

arXiv:2606.00995v1 Announce Type: new Abstract: Subliminal learning refers to a student language model acquiring a teacher's traits (e.g. a system-prompted preference for owls) when fine-tuned on the

Suppressing Forgery-Specific Shortcuts for Generalizable Deepfake Detection

Model ReleasesDGX agent

arXiv:2606.01843v1 Announce Type: cross Abstract: Deepfake detection suffers from poor generalization across forgery methods, as existing models tend to rely on spurious method-specific shortcuts that

SUPREME: A Multi-GPU Framework for Reproducible Image Unlearning Method Evaluation

HardwareDGX agent

arXiv:2606.00380v1 Announce Type: cross Abstract: Machine unlearning removes the influence of specific training data from a trained model without retraining it from scratch. Evaluating an unlearning m

Symbolic Neural Generation with Applications to Lead Discovery in Drug Design

Model ReleasesDGX agent

arXiv:2510.23379v2 Announce Type: replace-cross Abstract: We investigate a relatively under-explored class of hybrid neurosymbolic models that integrate symbolic learning with neural reasoning to cons

Synthetic Data from Cross-Domain Events for Large-Scale Recommendation Systems

ResearchDGX agent

arXiv:2606.00282v1 Announce Type: cross Abstract: Large-scale recommendation systems operate across diverse domains, yet they face the challenges of data sparsity and noisy implicit feedback. Traditio

T-POP: Test-Time Personalization with Online Preference Feedback

SafetyDGX agent

arXiv:2509.24696v2 Announce Type: replace-cross Abstract: Personalizing large language models (LLMs) to individual user preferences is a critical step beyond generating generically helpful responses.

T1: Tool-integrated Verification for Test-time Compute Scaling in Small Language Models

Model ReleasesDGX agent

arXiv:2504.04718v2 Announce Type: replace-cross Abstract: Recent studies have demonstrated that test-time compute scaling effectively improves the performance of small language models (sLMs). However,

TabChange: Precise Attribute Changes in Tabular Data

ResearchDGX agent

arXiv:2606.00503v1 Announce Type: cross Abstract: Modifying an attribute in tabular data often introduces an unnatural instance by breaking its relationships with other attributes. The modified instan

Tackling the Root of Misinformation by Teaching Laypeople about Logical Fallacies via Socratic Questioning and Critical Argumentation

TutorialsDGX agent

arXiv:2606.01020v1 Announce Type: new Abstract: Identifying logical fallacies in everyday discourse is challenging for many people. This challenge is amplified in the era of Large Language Models (LLM

Taming System Complexity: Demystifying Software Engineering Agents in Diagnosing Linux Kernel Faults

Model ReleasesDGX agent

arXiv:2505.19489v2 Announce Type: replace Abstract: The Linux kernel is a critical system, serving as the foundation for numerous systems. Bugs in the Linux kernel can cause serious consequences, affe

TAPS: Target-Aware Prefix Tree Selection for Diffusion-Drafted Speculative Decoding

ResearchDGX agent

arXiv:2606.00487v1 Announce Type: new Abstract: Using a diffusion model for parallel drafting is a promising approach for speculative decoding. By predicting tokens at multiple future positions in a s

Task diversity produces systematic transfer but inhibits continual reinforcement learning

Model ReleasesDGX agent

arXiv:2606.00880v1 Announce Type: cross Abstract: Continual reinforcement learning aims to produce agents that learn not only to improve at their current tasks but also to adapt as task distributions

TCAR-Gen: Temporal Graph Retrieval with Evidence Fusion for Knowledge-Grounded Generation

Model ReleasesDGX agent

arXiv:2606.00029v1 Announce Type: cross Abstract: Retrieval-augmented generation systems struggle with temporal reasoning and evidence fusion when answering complex questions over historical criminal

TECCI: Tricky Edits of Collected and Curated Images

Model ReleasesDGX agent

arXiv:2606.01213v1 Announce Type: cross Abstract: Despite tremendous recent progress, current text-guided image editing methods still struggle with many aspects of editing involving instruction follow

TechGraphRAG: An Agentic Graph-Augmented RAG Framework for Technical Literature Reasoning

AgentsDGX agent

arXiv:2606.01613v1 Announce Type: cross Abstract: This paper presents an agentic retrieval-augmented generation (RAG) framework for domain-specific technical reasoning support, instantiated over a cur

Temporally-Aligned Evaluation for Audio-Driven Talking Head Generation

Model ReleasesDGX agent

arXiv:2606.01031v1 Announce Type: cross Abstract: Audio-driven talking-head generation has advanced rapidly, yet existing evaluation protocols mainly rely on frame-wise metrics that assume strict temp

TERRA: Task-Embedded Reasoning and Representation Architecture for Cross-Domain Applications

ResearchDGX agent

arXiv:2606.01520v1 Announce Type: new Abstract: A single action-conditioned latent predictive architecture can in principle be trained on the structured state of a driving scene, a robot workspace, or

Test-Time Training for Zero-Resource Dense Retrieval Reranking

ResearchDGX agent

arXiv:2606.01070v1 Announce Type: cross Abstract: Dense retrievers excel at first-stage candidate generation but lack effective reranking in zero-resource settings. Existing approaches face a fundamen

The Alignment Curse: Modality Alignment Supercharges Audio Attacks via Text Transfer

SafetyDGX agent

arXiv:2602.02557v2 Announce Type: replace-cross Abstract: Recent advances in end-to-end trained omni-models have substantially improved audio capabilities by strengthening text-audio modality alignmen

The Case for Model Science: Verify, Explore, Steer, Refine

Model ReleasesDGX agent

arXiv:2606.01189v1 Announce Type: new Abstract: We argue that the AI community is now ready to move beyond benchmarking and consolidate scattered efforts in model analysis into a systematic discipline

The Deterministic Horizon: When Extended Reasoning Fails and Tool Delegation Becomes Necessary

AgentsDGX agent

arXiv:2606.00376v1 Announce Type: new Abstract: Extended chain-of-thought reasoning can degrade performance on deterministic state-tracking tasks, not due to preference biases, but limits rooted in th

The Geometry of Grokking: Norm Minimization on the Zero-Loss Manifold

ResearchDGX agent

arXiv:2511.01938v3 Announce Type: replace-cross Abstract: Grokking is a puzzling phenomenon in neural networks where full generalization occurs only after a substantial delay following the complete me

The Image Reconstruction Game: Drawing Common Ground Through Iterative Multimodal Dialogue

Model ReleasesDGX agent

arXiv:2606.01901v1 Announce Type: cross Abstract: We introduce the Image Reconstruction Game, a fully automated benchmark in which a vision-language model issues corrective instructions to an image ge

The New Social Image: How AI Competency and AI Proactivity Influence Self- and Peer-Perceptions in the Workplace

ResearchDGX agent

arXiv:2606.00182v1 Announce Type: cross Abstract: Human-AI collaboration is considered the most promising way to incorporate AI in the workplace. What remains unexplored are the experiential consequen

The Paradox of Outcome Optimization: A Causal Information-Theoretic Bound on Reasoning Shortcuts in LLMs

SafetyDGX agent

arXiv:2606.00674v1 Announce Type: cross Abstract: Large Language Models (LLMs) aligned via outcome-based Reinforcement Learning (RL) frequently exhibit a critical failure mode: they achieve high perfo

The Refusal--Compliance Tradeoff: A Large-Scale Safety Behavior Audit of Large Language Models

Model ReleasesDGX agent

arXiv:2605.05427v2 Announce Type: replace Abstract: Refusal rates are a poor proxy for LLM safety, i.e., a model may over-refuse benign prompts while still complying with harmful ones. We audit both f

The Role of Ambiguity in Error Prediction via Uncertainty Quantification

ResearchDGX agent

arXiv:2606.02093v1 Announce Type: cross Abstract: The task of Error Prediction, namely predicting whether a model output is correct, is commonly tackled with Uncertainty Quantification (UQ). However,

The Shape of Wisdom: Decision Trajectories in Language Models

Model ReleasesDGX agent

arXiv:2606.01202v1 Announce Type: new Abstract: Language models do not simply choose an answer at the output layer. In a 9,000-trajectory MMLU study across Qwen2.5-7B-Instruct, Llama-3.1-8B-Instruct,

ThinkSwitch: Context Distillation with LoRA and Weight Interpolation for Specific-Purpose Reasoning Tasks

ResearchDGX agent

arXiv:2606.01080v1 Announce Type: cross Abstract: Large language models often improve on difficult tasks by spending inference-time compute on a reasoning trace before producing the final answer. That

THRD: A Training-Free Multi-Turn Defense Framework for Jailbreak Attacks on Large Language Models

SafetyDGX agent

arXiv:2606.01738v1 Announce Type: cross Abstract: Multi-turn jailbreak attacks pose a growing threat to LLMs by exploiting conversational dynamics such as gradual escalation and cross-turn coordinatio

Threshold-Based Exclusive Batching for LLM Inference

HardwareDGX agent

arXiv:2606.00516v1 Announce Type: new Abstract: Mixed batching (MB)--interleaving prefill and decode in a single batch--has become the standard scheduling strategy for large language model (LLM) infer

TIGER: Traceable Inference with Graph-Based Evidence Routing for Mitigating Hallucinations in Multimodal Generation

Local AiDGX agent

arXiv:2606.00232v1 Announce Type: new Abstract: We study fact-level repair for multimodal generation, where a fluent output may contain specific facts that are not supported by the input. Existing inf

Time-Aware Diffusion based on Preference Disentanglement for Generative Recommendation

ApplicationsDGX agent

arXiv:2606.01670v1 Announce Type: cross Abstract: Recently, Generative Recommenders (GRs) have emerged as a transformative recommendation paradigm by replacing traditional item IDs with semantic indic

TimeSage-MT: A Multi-Turn Benchmark for Evaluating Agentic Time Series Reasoning

Model ReleasesDGX agent

arXiv:2606.01498v1 Announce Type: cross Abstract: Time series data inform critical decisions across many real-world domains. While large language model (LLM) agents can analyze data through natural la

TN-SHAP-G: Graph-Structured Tensor Network Surrogates for Shapley Values and Interactions

ResearchDGX agent

arXiv:2606.01540v1 Announce Type: cross Abstract: Shapley values are a widely used tool for attributing importance and interactions among input variables in black-box models, but their computation inv

Token Predictors Are Not Planners: Building Physically Grounded Causal Reasoners

Model ReleasesDGX agent

arXiv:2606.01810v1 Announce Type: new Abstract: Current benchmarks for embodied vision-language planning often favor linguistic next-token prediction over physically grounded next-state reasoning. Thi

ToolSelf: Unifying Task Execution and Self-Reconfiguration via Tool-Driven Emergent Adaptation

SafetyDGX agent

arXiv:2602.07883v3 Announce Type: replace Abstract: LLM-powered agentic systems excel at complex long-horizon tasks, but remain constrained by static configurations fixed before execution. Such rigidi

Topological Ignorability for Structural Causal Effects Beyond Means

Model ReleasesDGX agent

arXiv:2606.01184v1 Announce Type: cross Abstract: Many interventions alter the structure of an outcome distribution rather than its mean: they can split a population into disconnected regimes, create

Topological texture analysis of microscopy images of dynamic casein gelation and its relation to rheological properties

ResearchDGX agent

arXiv:2606.02048v1 Announce Type: new Abstract: We propose a novel computational toolbox that integrates Topological Data Analysis (TDA), Differential Box Counting (DBC), Multifractal Partition (MFP),

Toward accurate RUL and SoH estimation using reinforced graph-based physics-informed neural networks enhanced with dynamic weights

Model ReleasesDGX agent

arXiv:2507.09766v2 Announce Type: replace-cross Abstract: Accurate estimation of Remaining Useful Life (RUL) and State of Health (SoH) is essential for reliable Prognostics and Health Management (PHM)

Toward Robust In-Context Learning: Leveraging Out-of-distribution Proxies for Target Inaccessible Demonstration Retrieval

TutorialsDGX agent

arXiv:2606.00014v1 Announce Type: cross Abstract: Although studies have demonstrated that Large Language Models (LLMs) can perform well on Out-of-Distribution (OOD) tasks, their advantage tends to dim

Towards 3D-Aware Video Diffusion Models: Render-Free Human Motion Control with Mesh Tokenization

ResearchDGX agent

arXiv:2606.02000v1 Announce Type: cross Abstract: Diffusion models have shown remarkable success in video generation. However, whether such models are truly aware of the 3D structure underlying visual

Towards a General Intelligence and Interface for Wearable Health Data

AgentsDGX agent

arXiv:2605.22759v2 Announce Type: replace Abstract: While ubiquitous wearable sensors capture a wealth of behavioral and physiological information, effectively transforming these signals into personal

Towards a Physics Foundation Model

TutorialsDGX agent

arXiv:2509.13805v4 Announce Type: replace-cross Abstract: Foundation models have revolutionized natural language processing through a ``train once, deploy anywhere'' paradigm, where a single pre-train

Towards Resolving Optimization Conflicts Between Image- and Text-Based Person Re-Identification

ResearchDGX agent

arXiv:2606.02242v1 Announce Type: cross Abstract: The joint optimization of image-based (I2I) and text-based (T2I) person re-identification (ReID) is hindered by modality discrepancies and conflicting

Towards Understanding Modality Interaction in Multimodal Language Models via Partial Information Decomposition

SafetyDGX agent

arXiv:2606.00959v1 Announce Type: new Abstract: Understanding modality interaction in multimodal large language models (MLLMs) is central to reliable deployment. We introduce Partial Information Decom

TRACE: Trajectory Risk-Aware Compression for Long-Horizon Agent Safety

SafetyDGX agent

arXiv:2606.00611v1 Announce Type: new Abstract: Long-horizon LLM agents produce safety evidence across long trajectories, where sparse, delayed, and compositional risk signals often escape local moder

Tracing GenAI Literacy: Uncovering Student-AI Interaction Patterns in Academic Writing through Epistemic Network Analysis

ApplicationsDGX agent

arXiv:2606.00040v1 Announce Type: cross Abstract: As Generative AI (GenAI) becomes integral to education, fostering GenAI literacy is critical. However, current assessments largely rely on self-report

← Previous
1…180181182183184…358
Next →