AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
Human
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,851 results
21 Apr 2026

Downgrade to Upgrade: Optimizer Simplification Enhances Robustness in LLM Unlearning

SafetyDGX agent

arXiv:2510.00761v5 Announce Type: replace Abstract: Large language model (LLM) unlearning aims to surgically remove the influence of undesired data or knowledge from an existing model while preserving

DR-SAC: Distributionally Robust Soft Actor-Critic for Reinforcement Learning under Uncertainty

SafetyDGX agent

arXiv:2506.12622v2 Announce Type: replace Abstract: Deep reinforcement learning (RL) has achieved remarkable success, yet its deployment in real-world scenarios is often limited by vulnerability to en

DREAM: Dynamic Retinal Enhancement with Adaptive Multi-modal Fusion for Expert Precision Medical Report Generation

Model ReleasesDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.17209v1 Announce Type: new Abstract: Automating medical reports for retinal images requires a sophisticated blend of visual pattern recognition and deep clinical knowledge. Current Large Vi

DreamShot: Personalized Storyboard Synthesis with Video Diffusion Prior

SafetyDGX agent

arXiv:2604.17195v1 Announce Type: new Abstract: Storyboard synthesis plays a crucial role in visual storytelling, aiming to generate coherent shot sequences that visually narrate cinematic events with

DriveAgent-R1: Advancing VLM-based Autonomous Driving with Active Perception and Hybrid Thinking

Model ReleasesDGX agent

arXiv:2507.20879v3 Announce Type: replace Abstract: The advent of Vision-Language Models (VLMs) has significantly advanced end-to-end autonomous driving, demonstrating powerful reasoning abilities for

Driving in Corner Case: A Real-World Adversarial Closed-Loop Evaluation Platform for End-to-End Autonomous Driving

SafetyDGX agent

arXiv:2512.16055v2 Announce Type: replace Abstract: Safety-critical corner cases, difficult to collect in the real world, are crucial for evaluating end-to-end autonomous driving. Adversarial interact

Driving risk emerges from the required two-dimensional joint evasive acceleration

SafetyDGX agent

arXiv:2604.17841v1 Announce Type: new Abstract: Most autonomous driving safety benchmarks use time-to-collision (TTC) to assess risk and guide safe behaviour. However, TTC-based methods treat risk as

DSA-CycleGAN: A Domain Shift Aware CycleGAN for Robust Multi-Stain Glomeruli Segmentation

ResearchDGX agent

arXiv:2604.18368v1 Announce Type: new Abstract: A key challenge in segmentation in digital histopathology is inter- and intra-stain variations as it reduces model performance. Labelling each stain is

DSH-Bench: A Difficulty- and Scenario-Aware Benchmark with Hierarchical Subject Taxonomy for Subject-Driven Text-to-Image Generation

Model ReleasesDGX agent

arXiv:2603.08090v2 Announce Type: replace Abstract: Significant progress has been achieved in subject-driven text-to-image (T2I) generation, which aims to synthesize new images depicting target subjec

Dual Alignment Between Language Model Layers and Human Sentence Processing

SafetyDGX agent

arXiv:2604.18563v1 Announce Type: new Abstract: A recent study (Kuribayashi et al., 2025) has shown that human sentence processing behavior, typically measured on syntactically unchallenging construct

Dual-Anchoring: Addressing State Drift in Vision-Language Navigation

AgentsDGX agent

arXiv:2604.17473v1 Announce Type: new Abstract: Vision-Language Navigation(VLN) requires an agent to navigate through 3D environments by following natural language instructions. While recent Video Lar

Dual-End Consistency Model

ResearchDGX agent

arXiv:2602.10764v2 Announce Type: replace Abstract: The slow iterative sampling nature remains a major bottleneck for the practical deployment of diffusion and flow-based generative models. While cons

Dual Strategies for Test-Time Adaptation

ResearchDGX agent

arXiv:2604.17542v1 Announce Type: new Abstract: Conventional test-time adaptation (TTA) approaches typically adapt the model using only a small fraction of test samples, often those with low-entropy p

Dual-stream Spatio-Temporal GCN-Transformer Network for 3D Human Pose Estimation

Model ReleasesDGX agent

arXiv:2604.17688v1 Announce Type: new Abstract: 3D human pose estimation is a classic and important research direction in the field of computer vision. In recent years, Transformer-based methods have

Duality for the Adversarial Total Variation

ResearchDGX agent

arXiv:2604.18540v1 Announce Type: cross Abstract: Adversarial training of binary classifiers can be reformulated as regularized risk minimization involving a nonlocal total variation. Building on this

DualToken: Towards Unifying Visual Understanding and Generation with Dual Visual Vocabularies

ResearchDGX agent

arXiv:2503.14324v3 Announce Type: replace-cross Abstract: The differing representation spaces required for visual understanding and generation pose a challenge in unifying them within the autoregressi

DuConTE: Dual-Granularity Text Encoder with Topology-Constrained Attention for Text-attributed Graphs

Model ReleasesDGX agent

arXiv:2604.17411v1 Announce Type: new Abstract: Text-attributed graphs integrate semantic information of node texts with topological structure, offering significant value in various applications such

DuQuant++: Fine-grained Rotation Enhances Microscaling FP4 Quantization

Model ReleasesDGX agent

arXiv:2604.17789v1 Announce Type: cross Abstract: The MXFP4 microscaling format, which partitions tensors into blocks of 32 elements sharing an E8M0 scaling factor, has emerged as a promising substrat

DVAR: Adversarial Multi-Agent Debate for Video Authenticity Detection

AgentsDGX agent

arXiv:2604.16987v1 Announce Type: new Abstract: The rapid evolution of video generation technologies poses a significant challenge to media forensics, as conventional detection methods often fail to g

Dynamic Emotion and Personality Profiling for Multimodal Deception Detection

SafetyDGX agent

arXiv:2604.17037v1 Announce Type: new Abstract: Deception detection is of great significance for ensuring information security and conducting public opinion analysis, with personality factors and emot

Dynamic Eraser for Guided Concept Erasure in Diffusion Models

ResearchDGX agent

arXiv:2604.16483v1 Announce Type: new Abstract: Concept erasure in Text-To-Image (T2I) diffusion models is vital for safe content generation, but existing inference-time methods face significant limit

Dynamic Risk Assessment by Bayesian Attack Graphs and Process Mining

ResearchDGX agent

arXiv:2604.18080v1 Announce Type: cross Abstract: While attack graphs are useful for identifying major cybersecurity threats affecting a system, they do not provide operational support for determining

Dynamic Visual-semantic Alignment for Zero-shot Learning with Ambiguous Labels

SafetyDGX agent

arXiv:2604.17710v1 Announce Type: new Abstract: Zero-shot learning (ZSL) aims to recognize unseen classes without visual instances. However, existing methods usually assume clean labels, overlooking r

DynaWeb: Model-Based Reinforcement Learning of Web Agents

SafetyDGX agent

arXiv:2601.22149v2 Announce Type: replace Abstract: The development of autonomous web agents, powered by Large Language Models (LLMs) and reinforcement learning (RL), represents a significant step tow

E2E-GMNER: End-to-End Generative Grounded Multimodal Named Entity Recognition

ResearchDGX agent

arXiv:2604.17319v1 Announce Type: cross Abstract: Grounded Multimodal Named Entity Recognition (GMNER) aims to jointly identify named entity mentions in text, predict their semantic types, and ground

E2E-WAVE: End-to-End Learned Waveform Generation for Underwater Video Multicasting

ResearchDGX agent

arXiv:2604.17047v1 Announce Type: cross Abstract: We present E2E-WAVE, the first end-to-end learned waveform generation system for underwater video multicasting. Acoustic channels exhibit 20--46% bit

E3VS-Bench: A Benchmark for Viewpoint-Dependent Active Perception in 3D Gaussian Splatting Scenes

Model ReleasesDGX agent

arXiv:2604.17969v1 Announce Type: new Abstract: Visual search in 3D environments requires embodied agents to actively explore their surroundings and acquire task-relevant evidence. However, existing v

EarthSight: A Distributed Framework for Low-Latency Satellite Intelligence

ResearchDGX agent

arXiv:2511.10834v3 Announce Type: replace Abstract: Low-latency delivery of satellite imagery is essential for time-critical applications such as disaster response, intelligence, and infrastructure mo

EAST: Early Action Prediction Sampling Strategy with Token Masking

ResearchDGX agent

arXiv:2604.18367v1 Announce Type: new Abstract: Early action prediction seeks to anticipate an action before it fully unfolds, but limited visual evidence makes this task especially challenging. We in

EasyVideoR1: Easier RL for Video Understanding

Model ReleasesDGX agent

arXiv:2604.16893v1 Announce Type: new Abstract: Reinforcement learning from verifiable rewards (RLVR) has demonstrated remarkable effectiveness in improving the reasoning capabilities of large languag

EchoChain: A Full-Duplex Benchmark for State-Update Reasoning Under Interruptions

Model ReleasesDGX agent

arXiv:2604.16456v1 Announce Type: new Abstract: Real-time voice assistants must revise task state when users interrupt mid-response, but existing spoken-dialog benchmarks largely evaluate turn-based i

eCP: Equivariant Conformal Prediction with pre-trained models

ResearchDGX agent

arXiv:2602.03986v2 Announce Type: replace Abstract: Conformal prediction, a post-hoc, distribution-free, finite-sample method of uncertainty quantification that offers formal coverage guarantees under

EdgeVTP: Exploration of Latency-efficient Trajectory Prediction for Edge-based Embedded Vision Applications

ResearchDGX agent

arXiv:2604.16783v1 Announce Type: new Abstract: Vehicle trajectory prediction is central to highway perception, but deployment on roadside edge devices necessitates bounded, deterministic end-to-end l

Edit Fidelity Field: Semantics-Aware Region Isolation for Training-Free Scene Text Editing

ApplicationsDGX agent

arXiv:2604.17500v1 Announce Type: new Abstract: Scene text editing (STE) has achieved remarkable progress in accurately rendering target text through diffusion-based methods. However, we identify a cr

EditVerse: Unifying Image and Video Editing and Generation with In-Context Learning

Model ReleasesDGX agent

arXiv:2509.20360v3 Announce Type: replace Abstract: Recent advances in foundation models highlight a clear trend toward unification and scaling, showing emergent capabilities across diverse domains. W

EduRABSA: An Education Review Dataset for Aspect-based Sentiment Analysis Tasks

TutorialsDGX agent

arXiv:2508.17008v2 Announce Type: replace Abstract: Every year, most educational institutions seek and receive an enormous volume of text feedback from students on courses, teaching, and overall exper

EEG-Based Emergency Braking Intensity Prediction Using Blind Source Separation

ResearchDGX agent

arXiv:2604.18220v1 Announce Type: cross Abstract: Electroencephalography (EEG) signals have been promising for long-term braking intensity prediction but are prone to various artifacts that limit thei

Efficient Diffusion Models under Nonconvex Equality and Inequality constraints via Landing

SafetyDGX agent

arXiv:2604.17838v1 Announce Type: new Abstract: Generative modeling within constrained sets is essential for scientific and engineering applications involving physical, geometric, or safety requiremen

Efficient Federated RLHF via Zeroth-Order Policy Optimization

SafetyDGX agent

arXiv:2604.17747v1 Announce Type: new Abstract: This paper considers reinforcement learning from human feedback in a federated learning setting with resource-constrained agents, such as edge devices.

Efficient Inference for Coupled Hidden Markov Models in Continuous Time and Discrete Space

ResearchDGX agent

arXiv:2510.12916v2 Announce Type: replace-cross Abstract: Systems of interacting continuous-time Markov chains are a powerful model class, but inference is typically intractable in high dimensional se

Efficient Low-Resource Language Adaptation via Multi-Source Dynamic Logit Fusion

ResearchDGX agent

arXiv:2604.18106v1 Announce Type: new Abstract: Adapting large language models (LLMs) to low-resource languages (LRLs) is constrained by the scarcity of task data and computational resources. Although

Efficient Task Adaptation in Large Language Models via Selective Parameter Optimization

Model ReleasesDGX agent

arXiv:2604.17051v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated excellent performance in general language understanding, generation and other tasks. However, when fine-t

Ego-InBetween: Generating Object State Transitions in Ego-Centric Videos

ResearchDGX agent

arXiv:2604.17749v1 Announce Type: new Abstract: Understanding physical transformation processes is crucial for both human cognition and artificial intelligence systems, particularly from an egocentric

EgoSound: Benchmarking Sound Understanding in Egocentric Videos

Model ReleasesDGX agent

arXiv:2602.14122v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have recently achieved remarkable progress in vision-language understanding. Yet, human perception is inher

EgoWalk: A Multimodal Dataset for Robot Navigation in the Wild

ApplicationsDGX agent

arXiv:2505.21282v2 Announce Type: replace Abstract: Data-driven navigation algorithms are critically dependent on large-scale, high-quality real-world data collection for successful training and robus

Eluder dimension: localise it!

ResearchDGX agent

arXiv:2601.09825v2 Announce Type: replace Abstract: We establish a lower bound on the eluder dimension of generalised linear model classes, showing that standard eluder dimension-based analysis cannot

Embedding Arithmetic: A Lightweight, Tuning-Free Framework for Post-hoc Bias Mitigation in Text-to-Image Models

Model ReleasesDGX agent

arXiv:2604.18167v1 Announce Type: new Abstract: Modern text-to-image (T2I) models amplify harmful societal biases, challenging their ethical deployment. We introduce an inference-time method that reli

EmbodiedHead: Real-Time Listening and Speaking Avatar for Conversational Agents

ResearchDGX agent

arXiv:2604.17211v1 Announce Type: new Abstract: We present EmbodiedHead, a speech-driven talking-head framework that equips LLMs with real-time visual avatars for conversation. A practical embodied av

EmbodiedLGR: Integrating Lightweight Graph Representation and Retrieval for Semantic-Spatial Memory in Robotic Agents

Model ReleasesDGX agent

arXiv:2604.18271v1 Announce Type: new Abstract: As the world of agentic artificial intelligence applied to robotics evolves, the need for agents capable of building and retrieving memories and observa

EmbodiTTA: Resource-Efficient Test-Time Adaptation for Embodied Visual Systems

ApplicationsDGX agent

arXiv:2505.00986v2 Announce Type: replace-cross Abstract: Continual Test-time adaptation (CTTA) continuously adapts the deployed model on every incoming batch of data. While achieving optimal accuracy

Emergency Stopping for Liquid-manipulating Robots

SafetyDGX agent

arXiv:2604.16667v1 Announce Type: new Abstract: Manipulating open liquid containers is challenging because liquids are highly sensitive to vessel accelerations and jerks. Although spill-free liquid ma

Emergent Misalignment via In-Context Learning: Narrow in-context examples can produce broadly misaligned LLMs

Model ReleasesDGX agent

arXiv:2510.11288v4 Announce Type: replace Abstract: Recent work has shown that narrow finetuning can produce broadly misaligned LLMs, a phenomenon termed emergent misalignment (EM). While concerning,

Emergent Structured Representations Support Flexible In-Context Inference in Large Language Models

ResearchDGX agent

arXiv:2602.07794v3 Announce Type: replace Abstract: Large language models (LLMs) exhibit emergent behaviors suggestive of human-like reasoning. While recent work has identified structured conceptual r

emg2speech: Synthesizing speech from electromyography using self-supervised speech models

ResearchDGX agent

arXiv:2510.23969v2 Announce Type: replace-cross Abstract: We present a neuromuscular speech interface that translates electromyographic (EMG) signals recorded from orofacial muscles during speech arti

Emotion Collider: Dual Hyperbolic Mirror Manifolds for Sentiment Recovery via Anti Emotion Reflection

ResearchDGX agent

arXiv:2602.16161v3 Announce Type: replace-cross Abstract: Emotional expression underpins natural communication and effective human-computer interaction. We present Emotion Collider (EC-Net), a hyperbo

EmoVerse: A MLLMs-Driven Emotion Representation Dataset for Interpretable Visual Emotion Analysis

ResearchDGX agent

arXiv:2511.12554v2 Announce Type: replace Abstract: Visual Emotion Analysis (VEA) aims to bridge the affective gap between visual content and human emotional responses. Despite its promise, progress i

Employing General-Purpose and Biomedical Large Language Models with Advanced Prompt Engineering for Pharmacoepidemiologic Study Design

Model ReleasesDGX agent

arXiv:2604.17988v1 Announce Type: new Abstract: Background: The potential of large language models (LLMs) to automate and support pharmacoepidemiologic study design is an emerging area of interest, ye

Empowering Multi-Turn Tool-Integrated Agentic Reasoning with Group Turn Policy Optimization

SafetyDGX agent

arXiv:2511.14846v2 Announce Type: replace-cross Abstract: Training Large Language Models (LLMs) for multi-turn Tool-Integrated Reasoning (TIR) - where models iteratively reason, generate code, and ver

Enabling AI ASICs for Zero Knowledge Proof

HardwareDGX agent

arXiv:2604.17808v1 Announce Type: cross Abstract: Zero-knowledge proof (ZKP) provers remain costly because multi-scalar multiplication (MSM) and number-theoretic transforms (NTTs) dominate runtime as

Enabling Stroke-Level Structural Analysis of Hieroglyphic Scripts without Language-Specific Priors

ResearchDGX agent

arXiv:2601.05508v2 Announce Type: replace-cross Abstract: Hieroglyphs, as logographic writing systems, encode rich semantic and cultural information within their internal structural composition. Yet,

← Previous
1…886887888889890…998
Next →