AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,171
  • Agents7,461
  • Applications5,337
  • Concepts5
  • Hardware1,806
  • Industry6,146
  • Local Ai4,871
  • Model Releases23,435
  • Research19,874
  • Safety13,191
  • Syntheses17
  • Tools1,673
  • Tutorials3,355

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,171
  • Agents7,461
  • Applications5,337
  • Concepts5
  • Hardware1,806
  • Industry6,146
  • Local Ai4,871
  • Model Releases23,435
  • Research19,874
  • Safety13,191
  • Syntheses17
  • Tools1,673
  • Tutorials3,355

Source
Human
87,171Total entries
1Added by human
87,170Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,006 results
22 May 2026

Accelerated Test-Time Scaling with Model-Free Speculative Sampling

ResearchDGX agent

arXiv:2506.04708v3 Announce Type: replace Abstract: Language models have demonstrated remarkable capabilities in reasoning tasks through test-time scaling techniques like best-of-N sampling and tree s

Accelerating Vision Foundation Models with Drop-in Depthwise Convolution

ResearchDGX agent

arXiv:2605.22132v1 Announce Type: new Abstract: Pretrained vision foundation models deliver strong performance across tasks with limited fine-tuning. However, their Vision Transformer (ViT) backbones

Access Paths for Efficient Ordering with Large Language Models

ResearchDGX agent

arXiv:2509.00303v3 Announce Type: replace-cross Abstract: In this work, we present the exttt{LLM ORDER BY} semantic operator as a logical abstraction and conduct a systematic study of its physical imp

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Action with Visual Primitives

ResearchDGX agent

arXiv:2605.22183v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have emerged as a promising paradigm for generalist robotic manipulation. A common design in current architectures m

AesFormer: Transform Everyday Photos into Beautiful Memories

Model ReleasesDGX agent

arXiv:2605.22126v1 Announce Type: new Abstract: In everyday photography, aesthetically appealing moments are often captured with structural flaws (e.g., composition, camera viewpoint, or pose) that ex

AgentCo-op: Retrieval-Based Synthesis of Interoperable Multi-Agent Workflows

Model ReleasesDGX agent

arXiv:2605.20425v1 Announce Type: new Abstract: Designing multi-agent workflows is especially difficult in open-ended scientific settings where tasks lack curated training sets, reliable scalar evalua

Agentic Agile-V: From Vibe Coding to Verified Engineering in Software and Hardware Development

AgentsDGX agent

arXiv:2605.20456v1 Announce Type: cross Abstract: Agentic AI coding systems can inspect repositories, plan implementation steps, edit files, call tools, run tests, and submit pull requests. These capa

Agentic CLEAR: Automating Multi-Level Evaluation of LLM Agents

SafetyDGX agent

arXiv:2605.22608v1 Announce Type: new Abstract: Agentic systems are becoming more capable: agents define strategies, take actions, and interact with different environments. This autonomy poses serious

AgroTools: A Benchmark for Tool-Augmented Multimodal Agents in Agriculture

Model ReleasesDGX agent

arXiv:2605.22366v1 Announce Type: new Abstract: Agricultural decision-making increasingly requires multimodal systems that can transform visual observations into reliable, executable actions. However,

AgroVG: A Large-Scale Multi-Source Benchmark for Agricultural Visual Grounding

Model ReleasesDGX agent

arXiv:2605.22034v1 Announce Type: new Abstract: Visual grounding, the task of localizing objects described by natural-language expressions, is a foundational capability for agricultural AI systems, en

AlignPose: Generalizable 6D Pose Estimation via Multi-view Feature-metric Alignment

Model ReleasesDGX agent

arXiv:2512.20538v2 Announce Type: replace Abstract: Single-view RGB model-based object pose estimation methods achieve strong generalization but are fundamentally limited by depth ambiguity, clutter,

AMEL: Accumulated Message Effects on LLM Judgments

Model ReleasesDGX agent

arXiv:2605.22714v1 Announce Type: cross Abstract: Large language models are routinely used as automated evaluators: to review code, moderate content, or score outputs, often with many items passing th

Amplifying, Not Learning: Fine-Tuned AI Text Detectors Amplify a Pretrained Direction

SafetyDGX agent

arXiv:2605.21653v1 Announce Type: cross Abstract: AI text detectors amplify a pretrained typicality axis; they do not construct an AI-vs-human boundary. On raw encoders before any task supervision, pr

An Application-Layer Multi-Modal Covert-Channel Reference Monitor for LLM Agent Egress

AgentsDGX agent

arXiv:2605.20734v1 Announce Type: cross Abstract: A large language model (LLM) agent that sends messages can leak data inside them. Destination allowlists and content scanners do not police whether an

An Entity Linking Agent for Question Answering

AgentsDGX agent

arXiv:2508.03865v4 Announce Type: replace Abstract: Some Question Answering (QA) systems rely on knowledge bases (KBs) to provide accurate answers. Entity Linking (EL) plays a critical role in linking

An Evidence Hierarchy for Bayesian Object Classification via OSINT-Aided Heterogeneous Sensor Fusion

ResearchDGX agent

arXiv:2605.22259v1 Announce Type: cross Abstract: Heterogeneous sensor fusion is vital for detecting, localizing, and classifying CBRNE threats. However, individual sensors are often only capable of d

An Open Multi-Center Whole-Body FDG PET/CT Foundation Model for Tumor Segmentation

ResearchDGX agent

arXiv:2605.21835v1 Announce Type: cross Abstract: The synergistic interpretation of anatomical information from computed tomography (CT) and metabolic information from positron emission tomography (PE

Analytical and Experimental Force Analysis of a Soft Linear Pneumatic Actuator

ResearchDGX agent

arXiv:2605.21836v1 Announce Type: new Abstract: Soft sleeve actuators (SSAs) have recently been developed as a pneumatic actuation approach for wearable and assistive robotic systems. By integrating t

AnyMo: Geometry-Aware Setup-Agnostic Modeling of Human Motion in the Wild

TutorialsDGX agent

arXiv:2605.22715v1 Announce Type: cross Abstract: As wearable and mobile devices become increasingly embedded in daily life, they offer a practical way to continuously sense human motion in the wild.

ArabDiscrim: A Decade-Long Arabic Facebook Corpus on Racism and Discrimination

Model ReleasesDGX agent

arXiv:2605.22081v1 Announce Type: new Abstract: We present ArabDiscrim, a decade-long lexical resource and corpus of 293K public Arabic Facebook posts (2014--2024) discussing racism and discrimination

Artificial Intelligence Reshapes Microwave Photonics

AgentsDGX agent

arXiv:2605.21224v1 Announce Type: cross Abstract: As a rapidly emerging interdisciplinary field that intrinsically integrates microwave and photonics, microwave photonics (MWP) provides disruptive sol

Assisted Counterspeech Writing at the Crossroads of Hate Speech and Misinformation

ResearchDGX agent

arXiv:2605.22435v1 Announce Type: new Abstract: Hate speech and misinformation frequently co-occur online, amplifying prejudice and polarization. Given their scale, using Large Language Models (LLMs)

AtomicMotion: Learning Human Motion From Different Human Parts

Local AiDGX agent

arXiv:2605.22631v1 Announce Type: new Abstract: Accurately reconstructing full-body poses from sparse head and hand trajectories is a foundational challenge for immersive AR/VR telepresence. Current m

Attacking the Spike: On the Transferability and Security of Spiking Neural Networks to Adversarial Examples

ResearchDGX agent

arXiv:2209.03358v5 Announce Type: replace-cross Abstract: Spiking neural networks (SNNs) have attracted much attention for their high energy efficiency and recent advances in classification performanc

Auction-Consensus Algorithm with Learned Bidding Scheme for Multi-Robot Systems

Local AiDGX agent

arXiv:2605.21932v1 Announce Type: new Abstract: Multi-Robot Task Allocation (MRTA) is a central challenge in decentralized multi-agent systems, where teams of robots must cooperatively assign and exec

Audience Engagement with Arabic Women's Social Empowerment and Wellbeing: A Decadal Corpus

Model ReleasesDGX agent

arXiv:2605.22204v1 Announce Type: new Abstract: This paper presents the Arabic Women and Society Corpus, a ten year collection of 252,487 public Arabic Facebook posts related to women's empowerment an

AutoRPA: Efficient GUI Automation through LLM-Driven Code Synthesis from Interactions

AgentsDGX agent

arXiv:2605.21082v1 Announce Type: new Abstract: Large Language Model (LLM) based agents have demonstrated proficiency in multi-step interactions with graphical user interfaces (GUIs). While most resea

AVI-HT: Adaptive Vision-IMU Fusion for 3D Hand Tracking

ApplicationsDGX agent

arXiv:2605.21714v1 Announce Type: new Abstract: We present AVI-HT, an adaptive visual-IMU fusion approach for tracking 3D hand poses by jointly modeling the egocentric image with on-glove 6-DoF IMU si

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation

AgentsDGX agent

arXiv:2605.22816v1 Announce Type: cross Abstract: Vision-and-Language Navigation (VLN) requires an agent to ground language instructions to its own movement within a visual environment. While state-of

Balancing Uncertainty and Diversity of Samples: Leveraging Diversity of Least, High Confidence Samples for Effective Active Learning

TutorialsDGX agent

arXiv:2605.22169v1 Announce Type: new Abstract: Deep learning models, including Convolutional Neural Networks (CNNs) and Vision Transformers (ViTs), have achieved state-of-the-art performance on vario

BEiTScore: Reference-free Image Captioning Evaluation with an Efficient Cross-Encoder Model

Model ReleasesDGX agent

arXiv:2605.21728v1 Announce Type: cross Abstract: Image captioning evaluation remains a significant challenge, as vision-language models evolve toward more challenging capabilities such as generating

BeLink: Biomedical Entity Linking Meets Generative Re-Ranking

ApplicationsDGX agent

arXiv:2605.22501v1 Announce Type: new Abstract: Despite recent progress, Biomedical Entity Linking (BEL) with large language models (LLMs) remains computationally inefficient and challenging to deploy

Bernini: Latent Semantic Planning for Video Diffusion

ResearchDGX agent

arXiv:2605.22344v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) and diffusion models have each reached remarkable maturity: MLLMs excel at reasoning over heterogeneous multimo

Beyond Acoustic Emotion Recognition: Multimodal Pathos Analysis in Political Speech Using LLM-Based and Acoustic Emotion Models

Model ReleasesDGX agent

arXiv:2605.22732v1 Announce Type: cross Abstract: We investigate whether acoustic emotion recognition models can serve as proxies for the Pathos dimension in political speech analysis, as operationali

Beyond Benchmark Islands: Toward Representative Trustworthiness Evaluation for Agentic AI

Model ReleasesDGX agent

arXiv:2603.14987v2 Announce Type: replace Abstract: Agentic AI systems increasingly act through tool-augmented, multi-step workflows whose failures (unsafe tool use, unauthorised actions, social harm)

Beyond Chamfer Distance: Granular Order-aware Evaluation Metric For Online Mapping

Model ReleasesDGX agent

arXiv:2605.22578v1 Announce Type: new Abstract: Online map estimation is a crucial component of autonomous driving systems that reduces the reliance on costly high-definition maps. State-of-the-art (S

Beyond Euclidean Proximity: Repairing Latent World Models with Horizon-Matched Trajectory Reachability Metrics

Model ReleasesDGX agent

arXiv:2605.22164v1 Announce Type: cross Abstract: Latent world models can contain the state needed for control, yet their terminal-cost interface can expose the planner to the wrong decision-relevant

Beyond Pixels: Learning Invariant Rewards for Real-World Robotics From a Few Demonstrations

SafetyDGX agent

arXiv:2605.22123v1 Announce Type: new Abstract: Designing reward functions that generalize beyond controlled laboratory settings remains a fundamental challenge in reinforcement learning for robotics.

Beyond Temperature: Hyperfitting as a Late-Stage Geometric Expansion

Model ReleasesDGX agent

arXiv:2605.22579v1 Announce Type: new Abstract: Recent work has identified a counterintuitive phenomenon termed 'Hyperfitting', where fine-tuning Large Language Models (LLMs) to near-zero training los

Blind Spots in the Guard: How Domain-Camouflaged Injection Attacks Evade Detection in Multi-Agent LLM Systems

Model ReleasesDGX agent

arXiv:2605.22001v1 Announce Type: cross Abstract: Injection detectors deployed to protect LLM agents are calibrated on static, template-based payloads that announce themselves as override directives.

BodyReLux: Temporally Consistent Full-Body Video Relighting

ApplicationsDGX agent

arXiv:2605.21766v1 Announce Type: new Abstract: Being able to relight human performance is a fundamental task for post production and content creation. We present BodyReLux, a subject-specific video d

Boiling the Frog: A Multi-Turn Benchmark for Agentic Safety

Model ReleasesDGX agent

arXiv:2605.22643v1 Announce Type: new Abstract: Background. Traditional safety benchmarks for language models evaluate generated text: whether a model outputs toxic language, reproduces bias, or follo

Boundary-targeted Membership Inference Attacks on Safety Classifiers

Local AiDGX agent

arXiv:2605.22373v1 Announce Type: cross Abstract: Safety classifiers are essential safeguards within generative AI systems, filtering harmful content or identifying at-risk users when interacting with

Bounding-Box Trajectories Matter for Video Anomaly Detection

SafetyDGX agent

arXiv:2605.21957v1 Announce Type: new Abstract: Video anomaly detection is critical for public safety and security, yet remains highly challenging despite extensive research due to large variations in

Branch-Stochastic Model Predictive Control for Motion Planning under Multi-Modal Uncertainty with Scenario Clustering

SafetyDGX agent

arXiv:2605.22600v1 Announce Type: new Abstract: Motion planning for autonomous driving must account for multi-modal uncertainty in both the intentions and trajectories of surrounding vehicles. Handlin

Broadening Access to Transportation Safety Data with Generative AI: A Schema-Grounded Framework for Spatial Natural Language Queries

Local AiDGX agent

arXiv:2605.21712v1 Announce Type: new Abstract: Transportation safety analysis requires integrating crash records, roadway attributes, and geospatial data through GIS-based workflows, but access remai

Broken Memories: Detecting and Mitigating Memorization in Diffusion Models with Degraded Generations

ResearchDGX agent

arXiv:2605.22050v1 Announce Type: new Abstract: While diffusion models excel at generating high-quality images, their tendency to memorize training data poses significant privacy and copyright risks.

Cambrian-P: Pose-Grounded Video Understanding

ResearchDGX agent

arXiv:2605.22819v1 Announce Type: new Abstract: Camera pose matters. The position and orientation of each viewpoint define a shared spatial coordinate frame that relates observations across video fram

Can We Build a Monolithic Model for Fake Image Detection? SICA: Semantic-Induced Constrained Adaptation for Unified-Yet-Discriminative Artifact Feature Space Reconstruction

ApplicationsDGX agent

arXiv:2602.06676v4 Announce Type: replace Abstract: Fake Image Detection (FID), aiming at unified detection across four image forensic subdomains, is critical in real-world forensic scenarios. Compare

Case-Aware Medical Image Classification with Multimodal Knowledge Graphs and Reliability-Guided Refinement

SafetyDGX agent

arXiv:2605.22547v1 Announce Type: new Abstract: Deep learning has brought significant progress to medical image classification, yet most existing methods still rely on isolated visual evidence and can

Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactive Video Generation

ResearchDGX agent

arXiv:2602.02214v3 Announce Type: replace Abstract: To achieve real-time interactive video generation, current methods distill pretrained bidirectional video diffusion models into few-step autoregress

Causal Past Logic for Runtime Verification of Distributed LLM Agent Workflows

AgentsDGX agent

arXiv:2605.20923v1 Announce Type: cross Abstract: Distributed LLM agent workflows should not be monitored as if they produced a single sequential log. In an asynchronous execution, a decision can only

Cell Phantom Video Generation in Elliptical Fourier Descriptor Domain

ResearchDGX agent

arXiv:2605.22563v1 Announce Type: new Abstract: Training Deep Neural Networks for tracking individual cells in biomedical videos requires a large amount of annotated data. The annotation of videos for

Chain-of-thought obfuscation learned from output supervision can generalise to unseen tasks

TutorialsDGX agent

arXiv:2601.23086v2 Announce Type: replace Abstract: Chain-of-thought (CoT) reasoning provides a significant performance uplift to LLMs by enabling planning, exploration, and deliberation of their acti

Check Your LLM's Secret Dictionary! Five Lines of Code Reveal What Your LLM Learned (Including What It Shouldn't Have)

Model ReleasesDGX agent

arXiv:2605.22005v1 Announce Type: cross Abstract: We show that singular value decomposition of the lm_head} weight matrix of a transformer-based large language model -- requiring only five lines of Py

Chinese sensorimotor and embodiment norms for 3,000 lexicalized concepts

ResearchDGX agent

arXiv:2605.22616v1 Announce Type: new Abstract: Understanding how conceptual knowledge is grounded in bodily experience, and to what extent machine systems can acquire such knowledge without direct se

ChronoMedKG: A Temporally-Grounded Biomedical Knowledge Graph and Benchmark for Clinical Reasoning

Model ReleasesDGX agent

arXiv:2605.22734v1 Announce Type: new Abstract: Biomedical knowledge graphs (KGs) treat disease associations as static facts, but temporal information is crucial for clinical reasoning, e.g., a sympto

Circle-RoPE: Cone-like Decoupled Rotary Positional Embedding for Large Vision-Language Models

SafetyDGX agent

arXiv:2505.16416v3 Announce Type: replace Abstract: Rotary Position Embedding (RoPE) is widely adopted in large language models, but when applied to vision-language models (VLMs) it couples text and i

Claim-Selective Certification for High-Risk Medical Retrieval-Augmented Generation

ResearchDGX agent

arXiv:2605.21949v1 Announce Type: new Abstract: Medical RAG systems in high-risk QA settings are often evaluated through a single answer-or-abstain decision, but mixed evidence may support one claim,

Closed-Loop Sim-to-Real Reinforcement Learning for Deformable Microfiber Shape Control

SafetyDGX agent

arXiv:2605.21688v1 Announce Type: new Abstract: Autonomous contact-based micromanipulation is challenging because surface and interfacial interactions at the microscale are difficult to model accurate

← Previous
1…617618619620621…1034
Next →