AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “papers”

GridTimelineEvolution
12,056 results
14 Apr 2026

SpotFormer: Multi-Scale Spatio-Temporal Transformer for Facial Expression Spotting

Local AiDGX agent

arXiv:2407.20799v3 Announce Type: replace Abstract: Facial expression spotting, identifying periods where facial expressions occur in a video, is a significant yet challenging task in facial expressio

Spotlight and Shadow: Attention-Guided Dual-Anchor Introspective Decoding for MLLM Hallucination Mitigation

TutorialsDGX agent

arXiv:2604.10071v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have demonstrated remarkable reasoning capabilities yet continue to suffer from hallucination, where generated

StableTTA: Training-Free Test-Time Adaptation that Improves Model Accuracy on ImageNet1K to 96%

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.04552v2 Announce Type: replace-cross Abstract: Ensemble methods are widely used to improve predictive performance, but their effectiveness often comes at the cost of increased memory usage

STGV: Spatio-Temporal Hash Encoding for Gaussian-based Video Representation

ResearchDGX agent

arXiv:2604.10910v1 Announce Type: new Abstract: 2D Gaussian Splatting (2DGS) has recently become a promising paradigm for high-quality video representation. However, existing methods employ content-ag

Stop Fixating on Prompts: Reasoning Hijacking and Constraint Tightening for Red-Teaming LLM Agents

SafetyDGX agent

arXiv:2604.05549v2 Announce Type: replace Abstract: With the widespread application of LLM-based agents across various domains, their complexity has introduced new security threats. Existing red-team

SynthAgent: Adapting Web Agents with Synthetic Supervision

ResearchDGX agent

arXiv:2511.06101v3 Announce Type: replace-cross Abstract: Web agents struggle to adapt to new websites due to the scarcity of environment specific tasks and demonstrations. Recent works have explored

TacMan-Turbo: Proactive Tactile Control for Robust and Efficient Articulated Object Manipulation

TutorialsDGX agent

arXiv:2508.02204v4 Announce Type: replace Abstract: Adept manipulation of articulated objects is essential for robots to operate successfully in human environments. Such manipulation requires both eff

TARAC: Mitigating Hallucination in LVLMs via Temporal Attention Real-time Accumulative Connection

ResearchDGX agent

arXiv:2504.04099v2 Announce Type: replace-cross Abstract: Large Vision-Language Models have demonstrated remarkable capabilities, yet they suffer from hallucinations that limit practical deployment. W

Tessera: Unlocking Heterogeneous GPUs through Kernel-Granularity Disaggregation

HardwareDGX agent

arXiv:2604.10180v1 Announce Type: cross Abstract: Disaggregation maps parts of an AI workload to different types of GPUs, offering a path to utilize modern heterogeneous GPU clusters. However, existin

The Impact of Federated Learning on Distributed Remote Sensing Archives

Local AiDGX agent

arXiv:2604.11562v1 Announce Type: new Abstract: Remote sensing archives are inherently distributed: Earth observation missions such as Sentinel-1, Sentinel-2, and Sentinel-3 have collectively accumula

The Phantom of PCIe: Constraining Generative Artificial Intelligences for Practical Peripherals Trace Synthesizing

ResearchDGX agent

arXiv:2411.06376v3 Announce Type: replace-cross Abstract: Peripheral Component Interconnect Express (PCIe) is the de facto interconnect standard for high-speed peripherals and CPUs. The development of

The Rise and Fall of G in AGI

Model ReleasesDGX agent

arXiv:2604.09911v1 Announce Type: cross Abstract: In the psychological literature the term `general intelligence' describes correlations between abilities and not simply the number of abilities. This

The Weight of a Bit: EMFI Sensitivity Analysis of Embedded Deep Learning Models

ResearchDGX agent

arXiv:2602.16309v2 Announce Type: replace-cross Abstract: Fault injection attacks on embedded neural network models have been shown as a potent threat. Numerous works studied resilience of models from

TokUR: Token-Level Uncertainty Estimation for Large Language Model Reasoning

ResearchDGX agent

arXiv:2505.11737v4 Announce Type: replace-cross Abstract: While Large Language Models (LLMs) have demonstrated impressive capabilities, their output quality remains inconsistent across various applica

TorchUMM: A Unified Multimodal Model Codebase for Evaluation, Analysis, and Post-training

Model ReleasesDGX agent

arXiv:2604.10784v1 Announce Type: new Abstract: Recent advances in unified multimodal models (UMMs) have led to a proliferation of architectures capable of understanding, generating, and editing acros

Towards an Appropriate Level of Reliance on AI: A Preliminary Reliance-Control Framework for AI in Software Engineering

ResearchDGX agent

arXiv:2604.10530v1 Announce Type: cross Abstract: How software developers interact with Artificial Intelligence (AI)-powered tools, including Large Language Models (LLMs), plays a vital role in how th

Towards Efficient Large Vision-Language Models: A Comprehensive Survey on Inference Strategies

ResearchDGX agent

arXiv:2603.27960v2 Announce Type: replace-cross Abstract: Although Large Vision Language Models (LVLMs) have demonstrated impressive multimodal reasoning capabilities, their scalability and deployment

Training Deep Visual Networks Beyond Loss and Accuracy Through a Dynamical Systems Approach

ResearchDGX agent

arXiv:2604.09716v1 Announce Type: cross Abstract: Deep visual recognition models are usually trained and evaluated using metrics such as loss and accuracy. While these measures show whether a model is

Training-Free Model Ensemble for Single-Image Super-Resolution via Strong-Branch Compensation

ResearchDGX agent

arXiv:2604.11564v1 Announce Type: new Abstract: Single-image super-resolution has progressed from deep convolutional baselines to stronger Transformer and state-space architectures, yet the correspond

Training-Free Multi-User Generative Semantic Communications via Null-Space Diffusion Sampling

ResearchDGX agent

arXiv:2405.09866v2 Announce Type: replace-cross Abstract: In recent years, novel communication strategies have emerged to face the challenges that the increased number of connected devices and the hig

Transformers Learn the Optimal DDPM Denoiser for Multi-Token GMMs

TutorialsDGX agent

arXiv:2604.10074v1 Announce Type: new Abstract: Transformer-based diffusion models have demonstrated remarkable performance at generating high-quality samples. However, our theoretical understanding o

Tuning Language Models for Robust Prediction of Diverse User Behaviors

ApplicationsDGX agent

arXiv:2505.17682v2 Announce Type: replace-cross Abstract: Predicting user behavior is essential for intelligent assistant services, yet deep learning models often struggle to capture long-tailed behav

UHD-GPGNet: UHD Video Denoising via Gaussian-Process-Guided Local Spatio-Temporal Modeling

Local AiDGX agent

arXiv:2604.11014v1 Announce Type: new Abstract: Ultra-high-definition (UHD) video denoising requires simultaneously suppressing complex spatio-temporal degradations, preserving fine textures and chrom

Ultra-Low-Dimensional Prompt Tuning via Random Projection

Model ReleasesDGX agent

arXiv:2502.04501v3 Announce Type: replace Abstract: Large language models achieve state-of-the-art performance but are increasingly costly to fine-tune. Prompt tuning is a parameter-efficient fine-tun

Unified Graph Prompt Learning via Low-Rank Graph Message Prompting

Model ReleasesDGX agent

arXiv:2604.11257v1 Announce Type: new Abstract: Graph Data Prompt (GDP), which introduces specific prompts in graph data for efficiently adapting pre-trained GNNs, has become a mainstream approach to

Unified Unsupervised and Sparsely-Supervised 3D Object Detection by Semantic Pseudo-Labeling and Prototype Learning

AgentsDGX agent

arXiv:2602.21484v2 Announce Type: replace Abstract: 3D object detection is essential for autonomous driving and robotic perception, yet its reliance on large-scale manually annotated data limits scala

Use of AI Tools: Guidelines to Maintain Academic Integrity in Computing Colleges

TutorialsDGX agent

arXiv:2604.11111v1 Announce Type: cross Abstract: The rapid adoption of AI tools such as ChatGPT has significantly transformed academic practices, offering considerable benefits for both students and

Using Unwrapped Full Color Space Palette Recording to Measure Exposedness of a Vehicle Exterior Parts for External Human Machine Interfaces

AgentsDGX agent

arXiv:2604.11406v1 Announce Type: new Abstract: One of the concerns with autonomous vehicles is their ability to communicate their intent to other road users, specially pedestrians, in order to preven

Variable Selection Using Relative Importance Rankings

Model ReleasesDGX agent

arXiv:2509.10853v2 Announce Type: replace-cross Abstract: Although conceptually related, variable selection and relative importance (RI) analysis have been treated quite differently in the literature.

VERTIGO: Visual Preference Optimization for Cinematic Camera Trajectory Generation

ResearchDGX agent

arXiv:2604.02467v2 Announce Type: replace-cross Abstract: Cinematic camera control relies on a tight feedback loop between director and cinematographer, where camera motion and framing are continuousl

Vestibular reservoir computing

ResearchDGX agent

arXiv:2604.09943v1 Announce Type: new Abstract: Reservoir computing (RC) is a computational framework known for its training efficiency, making it ideal for physical hardware implementations. However,

Vibe-driven model-based engineering

ResearchDGX agent

arXiv:2604.10645v1 Announce Type: cross Abstract: There is a pressing need for better development methods and tools to keep up with the growing demand and increasing complexity of new software systems

VisText-Mosquito: A Unified Multimodal Dataset for Visual Detection, Segmentation, and Textual Explanation on Mosquito Breeding Sites

ResearchDGX agent

arXiv:2506.14629v3 Announce Type: replace-cross Abstract: Mosquito-borne diseases pose a major global health risk, requiring early detection and proactive control of breeding sites to prevent outbreak

Visual Late Chunking: An Empirical Study of Contextual Chunking for Efficient Visual Document Retrieval

ResearchDGX agent

arXiv:2604.10167v1 Announce Type: cross Abstract: Multi-vector models dominate Visual Document Retrieval (VDR) due to their fine-grained matching capabilities, but their high storage and computational

VLMaterial: Vision-Language Model-Based Camera-Radar Fusion for Physics-Grounded Material Identification

SafetyDGX agent

arXiv:2604.11671v1 Announce Type: cross Abstract: Accurate material recognition is a fundamental capability for intelligent perception systems to interact safely and effectively with the physical worl

WARPED: Wrist-Aligned Rendering for Robot Policy Learning from Egocentric Human Demonstrations

SafetyDGX agent

arXiv:2604.10809v1 Announce Type: new Abstract: Recent advancements in learning from human demonstration have shown promising results in addressing the scalability and high cost of data collection req

WBCBench 2026: A Challenge for Robust White Blood Cell Classification Under Class Imbalance

Model ReleasesDGX agent

arXiv:2604.10797v1 Announce Type: new Abstract: We present WBCBench 2026, an ISBI challenge and benchmark for automated WBC classification designed to stress-test algorithms under three key difficulti

What Do Vision-Language Models Encode for Personalized Image Aesthetics Assessment?

ApplicationsDGX agent

arXiv:2604.11374v1 Announce Type: cross Abstract: Personalized image aesthetics assessment (PIAA) is an important research problem with practical real-world applications. While methods based on vision

When Can You Poison Rewards? A Tight Characterization of Reward Poisoning in Linear MDPs

SafetyDGX agent

arXiv:2604.10062v1 Announce Type: new Abstract: We study reward poisoning attacks in reinforcement learning (RL), where an adversary manipulates rewards within constrained budgets to force the target

Who Gets Which Message? Auditing Demographic Bias in LLM-Generated Targeted Text

Model ReleasesDGX agent

arXiv:2601.17172v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly capable of generating personalized, persuasive text at scale, raising new questions about bias a

Who Handles Orientation? Investigating Invariance in Feature Matching

TutorialsDGX agent

arXiv:2604.11809v1 Announce Type: new Abstract: Finding matching keypoints between images is a core problem in 3D computer vision. However, modern matchers struggle with large in-plane rotations. A st

Why Supervised Fine-Tuning Fails to Learn: A Systematic Study of Incomplete Learning in Large Language Models

Model ReleasesDGX agent

arXiv:2604.10079v1 Announce Type: new Abstract: Supervised Fine-Tuning (SFT) is the standard approach for adapting large language models (LLMs) to downstream tasks. However, we observe a persistent fa

WM-DAgger: Enabling Efficient Data Aggregation for Imitation Learning with World Models

SafetyDGX agent

arXiv:2604.11351v1 Announce Type: new Abstract: Imitation learning is a powerful paradigm for training robotic policies, yet its performance is limited by compounding errors: minor policy inaccuracies

Wolkowicz-Styan Upper Bound on the Hessian Eigenspectrum for Cross-Entropy Loss in Nonlinear Smooth Neural Networks

ResearchDGX agent

arXiv:2604.10202v1 Announce Type: cross Abstract: Neural networks (NNs) are central to modern machine learning and achieve state-of-the-art results in many applications. However, the relationship betw

YIELD: A Large-Scale Dataset and Evaluation Framework for Information Elicitation Agents

SafetyDGX agent

arXiv:2604.10968v1 Announce Type: new Abstract: Most conversational agents (CAs) are designed to satisfy user needs through user-driven interactions. However, many real-world settings, such as academi

ZoomR: Memory Efficient Reasoning through Multi-Granularity Key Value Retrieval

SafetyDGX agent

arXiv:2604.10898v1 Announce Type: new Abstract: Large language models (LLMs) have shown great performance on complex reasoning tasks but often require generating long intermediate thoughts before reac

13 Apr 2026

2D or 3D: Who Governs Salience in VLA Models? -- Tri-Stage Token Pruning Framework with Modality Salience Awareness

ResearchDGX agent

arXiv:2604.09244v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have emerged as the mainstream of embodied intelligence. Recent VLA models have expanded their input modalities fr

Adaptive Planning for Multi-Attribute Controllable Summarization with Monte Carlo Tree Search

Model ReleasesDGX agent

arXiv:2509.26435v2 Announce Type: replace-cross Abstract: Controllable summarization moves beyond generic outputs toward human-aligned summaries guided by specified attributes. In practice, the interd

Adaptive Rigor in AI System Evaluation using Temperature-Controlled Verdict Aggregation via Generalized Power Mean

Model ReleasesDGX agent

arXiv:2604.08595v1 Announce Type: cross Abstract: Existing evaluation methods for LLM-based AI systems, such as LLM-as-a-Judge, verdict systems, and NLI, do not always align well with human assessment

Adaptive Tuning of Parameterized Traffic Controllers via Multi-Agent Reinforcement Learning

Model ReleasesDGX agent

arXiv:2512.07417v2 Announce Type: replace Abstract: Effective traffic control is essential for mitigating congestion in transportation networks. Conventional traffic management strategies, including r

AgentSociety: Large-Scale Simulation of LLM-Driven Generative Agents Advances Understanding of Human Behaviors and Society

SafetyDGX agent

arXiv:2502.08691v2 Announce Type: replace-cross Abstract: Understanding human behavior and society is a central focus in social sciences, with the rise of generative social science marking a significa

Anchored Sliding Window: Toward Robust and Imperceptible Linguistic Steganography

Model ReleasesDGX agent

arXiv:2604.09066v1 Announce Type: new Abstract: Linguistic steganography based on language models typically assumes that steganographic texts are transmitted without alteration, making them fragile to

Attention-Based Sampler for Diffusion Language Models

ResearchDGX agent

arXiv:2604.08564v1 Announce Type: new Abstract: Auto-regressive models (ARMs) have established a dominant paradigm in language modeling. However, their strictly sequential decoding paradigm imposes fu

B-MoE: A Body-Part-Aware Mixture-of-Experts 'All Parts Matter' Approach to Micro-Action Recognition

ResearchDGX agent

arXiv:2603.24245v3 Announce Type: replace Abstract: Micro-actions, fleeting and low-amplitude motions, such as glances, nods, or minor posture shifts, carry rich social meaning but remain difficult fo

Balancing User Preferences by Social Networks: A Condition-Guided Social Recommendation Model for Mitigating Popularity Bias

SafetyDGX agent

arXiv:2405.16772v2 Announce Type: replace-cross Abstract: Social recommendation models weave social interactions into their design to provide uniquely personalized recommendation results for users. Ho

Beyond Flicker: Detecting Kinematic Inconsistencies for Generalizable Deepfake Video Detection

ResearchDGX agent

arXiv:2512.04175v2 Announce Type: replace Abstract: Generalizing deepfake detection to unseen manipulations remains a key challenge. A recent approach to tackle this issue is to train a network with p

Boosting Brain-inspired Path Integration Efficiency via Learning-based Replication of Continuous Attractor Neurodynamics

Model ReleasesDGX agent

arXiv:2511.17687v2 Announce Type: replace Abstract: The brain's Path Integration (PI) mechanism offers substantial guidance and inspiration for Brain-Inspired Navigation (BIN). However, the PI capabil

Bridging SFT and RL: Dynamic Policy Optimization for Robust Reasoning

SafetyDGX agent

arXiv:2604.08926v1 Announce Type: new Abstract: Post-training paradigms for Large Language Models (LLMs), primarily Supervised Fine-Tuning (SFT) and Reinforcement Learning (RL), face a fundamental dil

Building Better Environments for Autonomous Cyber Defence

AgentsDGX agent

arXiv:2604.08805v1 Announce Type: cross Abstract: In November 2025, the authors ran a workshop on the topic of what makes a good reinforcement learning (RL) environment for autonomous cyber defence (A

CaRLi-V: Camera-RADAR-LiDAR Point-Wise 3D Velocity Estimation

ResearchDGX agent

arXiv:2511.01383v2 Announce Type: replace Abstract: Accurate point-wise velocity estimation in 3D is crucial for robot interaction with non-rigid dynamic agents, enabling robust performance in path pl

← Previous
1…195196197198199…201
Next →