AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,171
  • Agents7,461
  • Applications5,337
  • Concepts5
  • Hardware1,806
  • Industry6,146
  • Local Ai4,871
  • Model Releases23,435
  • Research19,874
  • Safety13,191
  • Syntheses17
  • Tools1,673
  • Tutorials3,355

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,171
  • Agents7,461
  • Applications5,337
  • Concepts5
  • Hardware1,806
  • Industry6,146
  • Local Ai4,871
  • Model Releases23,435
  • Research19,874
  • Safety13,191
  • Syntheses17
  • Tools1,673
  • Tutorials3,355

Source
Human
87,171Total entries
1Added by human
87,170Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,006 results
22 May 2026

MTR-Bench: A Comprehensive Benchmark for Multi-Turn Reasoning Evaluation

Model ReleasesDGX agent

arXiv:2505.17123v3 Announce Type: replace Abstract: Recent advances in Large Language Models (LLMs) have shown promising results in complex reasoning tasks. However, current evaluations predominantly

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering

ResearchDGX agent

arXiv:2605.22269v1 Announce Type: new Abstract: Long streaming video QA remains challenging due to growing visual tokens and limited reasoning length of large language models (LLMs). KV-caching stores

Multi-scale interaction network for stereo image super-resolution

ResearchDGX agent

arXiv:2605.21913v1 Announce Type: new Abstract: Stereo image super-resolution aims to generate high-resolution images by leveraging complementary information from binocular systems. Although previous

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Multi-Stage Training for Abusive Comment Detection in Indic Languages

ResearchDGX agent

arXiv:2605.22380v1 Announce Type: new Abstract: In recent years social media has become an increasingly popular tool for communication. People use it to share their ideas, exchange information, and di

N3P: Accelerated Automated Parking via a Learning-Based Naturalistic Three-Stage Scheme

AgentsDGX agent

arXiv:2605.22722v1 Announce Type: new Abstract: Autonomous parking requires efficient path planning that ensures kinematic feasibility and collision avoidance in constrained environments. Hybrid A* is

NaviAgent: Graph-Driven Bilevel Planning for Scalable Tool Orchestration

SafetyDGX agent

arXiv:2506.19500v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) increasingly act as function-call agents that invoke external tools to tackle tasks beyond their static knowledge

Network-Based Interventions for HIV Prevention via Cascade-Aware Suppression of Transmission

ApplicationsDGX agent

arXiv:2605.20218v1 Announce Type: cross Abstract: Treating and preventing Human Immunodeficiency Virus (HIV) remains a critical global health challenge. While antiretroviral therapy provides a path to

No Pose, No Problem in 4D: Feed-Forward Dynamic Gaussians from Unposed Multi-View Videos

ResearchDGX agent

arXiv:2605.22190v1 Announce Type: new Abstract: Recent feed-forward 3D gaussian splatting methods have made dramatic progress on individual aspects of 3D scene reconstruction, but no existing method j

Noise-Space Attribution and Control of Chunk-Boundary Artifact

SafetyDGX agent

arXiv:2603.11642v2 Announce Type: replace Abstract: Action chunking is widely used in generative visuomotor policies, yet the recurring execution discontinuities at chunk boundaries still lack a mecha

Non-Contact Vibration-Based Damage Detection of Civil Structures Using a Cost-Effective Autonomous UAV

SafetyDGX agent

arXiv:2605.21914v1 Announce Type: new Abstract: This paper presents a non-contact approach for vibration-based structural damage detection using an autonomous and customized cost-effective unmanned ae

Not All Starting Points Are Equal: Pre-trained Priors and Their Outsized Impact on Person Identification

Model ReleasesDGX agent

arXiv:2507.17640v3 Announce Type: replace Abstract: Recent years have seen an explosion of diverse general purpose pre-training methodologies for computer vision. However, the impact that these pre-tr

OCELOT: Odometry and Contact Estimation for Legged Robots

ResearchDGX agent

arXiv:2605.21863v1 Announce Type: new Abstract: One of the significant challenges in legged robotics is achieving accurate odometry using only onboard proprioceptive sensors. In this study, we present

On the Complexity of Entailment for Cumulative Propositional Dependence Logics

ResearchDGX agent

arXiv:2605.21113v1 Announce Type: cross Abstract: This paper establishes and proves complexity results for entailment for cumulative propositional dependence logic and for cumulative propositional log

One prompt is not enough: Instruction Sensitivity Undermines Embedding Model Evaluation

ResearchDGX agent

arXiv:2605.22544v1 Announce Type: new Abstract: Instruction embedding models have become common among state-of-the-art models, however are evaluated using a single prompt per task. The single-point ev

One Sentence, One Drama: Personalized Short-Form Drama Generation via Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2605.22144v1 Announce Type: new Abstract: Existing approaches for digital short-drama production typically rely on one-shot LLM generated scripts and loosely coupled pipelines, which fail to sat

Open Materials 2024 (OMat24) Inorganic Materials Dataset and Models

ResearchDGX agent

arXiv:2410.12771v2 Announce Type: replace-cross Abstract: The ability to discover new materials with desirable properties is critical for numerous applications from helping mitigate climate change to

Open-source LLMs administer maximum electric shocks in a Milgram-like obedience experiment

SafetyDGX agent

arXiv:2605.21401v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed as autonomous agents that make sequences of decisions over extended interactions in high-stakes

Open-World Evaluations for Measuring Frontier AI Capabilities

Model ReleasesDGX agent

arXiv:2605.20520v1 Announce Type: new Abstract: Benchmark-based evaluation remains important for tracking frontier AI progress. But it can both overstate and understate deployed capability because it

OPERA: An Agent for Image Restoration with End-to-End Joint Planning-Execution Optimization

AgentsDGX agent

arXiv:2605.22104v1 Announce Type: new Abstract: Real-world image restoration is challenging due to complex and interacting mixed degradations. Recent agent-based approaches address this problem by com

Optical Quantum Mixed-State Reconstruction With Multiple Deep Learning Approaches

ResearchDGX agent

arXiv:2407.01734v4 Announce Type: replace-cross Abstract: Quantum state tomography is a crucial technique for characterizing the state of a quantum system, which is essential for many applications in

Optimus: A Robust Defense Framework for Mitigating Toxicity while Fine-Tuning Conversational AI

SafetyDGX agent

arXiv:2507.05660v3 Announce Type: replace-cross Abstract: Customizing Large Language Models (LLMs) on untrusted datasets poses severe risks of injecting toxic behaviors. In this work, we introduce Opt

ORBIS: Output-Guided Token Reduction with Distribution-Aware Matching for Video Diffusion Acceleration

HardwareDGX agent

arXiv:2605.22015v1 Announce Type: new Abstract: Diffusion Transformer (DiT) has emerged as a powerful model architecture for generating high-quality images and videos. In the case of video DiT, 3D Spa

OSCToM: RL-Guided Adversarial Generation for High-Order Theory of Mind

Model ReleasesDGX agent

arXiv:2605.20423v1 Announce Type: new Abstract: Large Language Models (LLMs) perform well on many language tasks, but their Theory of Mind (ToM) reasoning is still uneven in complex social settings. E

OSS: Open Suturing Skills Vision-Based Assessment Challenge 2024-2025

Model ReleasesDGX agent

arXiv:2605.22200v1 Announce Type: new Abstract: Achieving high levels of surgical skill through effective training is essential for optimal patient outcomes. Automated, data-driven skill assessment ho

PALS: Power-Aware LLM Serving for Mixture-of-Experts Models

HardwareDGX agent

arXiv:2605.21427v1 Announce Type: new Abstract: Large language model (LLM) inference has become a dominant workload in modern data centers, driving significant GPU utilization and energy consumption.

Parallel OctoMapping: A Scalable Framework for Enhanced Path Planning in Autonomous Navigation

SafetyDGX agent

arXiv:2603.22508v2 Announce Type: replace Abstract: Mapping is essential in robotics and autonomous systems because it provides the spatial foundation for path planning. Efficient mapping enables plan

PartCo: Part-Level Correspondence Priors Enhance Category Discovery

Model ReleasesDGX agent

arXiv:2509.22769v2 Announce Type: replace Abstract: Generalized Category Discovery (GCD) aims to identify both known and novel categories within unlabeled data by leveraging a set of labeled examples

Pattern-and-root inflectional morphology: the Arabic broken plural

ResearchDGX agent

arXiv:2605.22310v1 Announce Type: new Abstract: We present a substantially implemented model of description of the inflectional morphology of Arabic nouns, with special attention to the management of

Perception or Prejudice: Can MLLMs Go Beyond First Impressions of Personality?

Model ReleasesDGX agent

arXiv:2605.22109v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) are increasingly deployed in human-facing roles where personality perception is critical, yet existing benchm

Personality Engineering with AI Agents: A New Methodology for Negotiation Research

TutorialsDGX agent

arXiv:2605.20554v1 Announce Type: new Abstract: According to canonical negotiation theory, people's success in a negotiation depends on how well they balance competing demands--empathizing and asserti

PGDG: Physically Grounded Data Generation for Robust Bimanual Policy Learning from a Single Demonstration

SafetyDGX agent

arXiv:2605.21710v1 Announce Type: new Abstract: Behavior cloning for contact-rich bimanual manipulation remains challenging because diverse demonstrations are expensive to collect, and even small dist

Physiology and Anatomy Aware Inverse Inference of Myocardial Infarction for Cardiac Digital Twin

Model ReleasesDGX agent

arXiv:2605.22044v1 Announce Type: new Abstract: Accurate localization of myocardial infarction is essential for risk stratification. While LGE-MRI remains the gold standard, it is resource-intensive.

PhysX-Omni: Unified Simulation-Ready Physical 3D Generation for Rigid, Deformable, and Articulated Objects

SafetyDGX agent

arXiv:2605.21572v1 Announce Type: new Abstract: Simulation-ready physical 3D assets have emerged as a promising direction owing to their broad applicability in downstream tasks. However, most existing

PIU: Proximity-guided Identity Unlearning in ID-Conditioned Diffusion Models

ResearchDGX agent

arXiv:2605.22311v1 Announce Type: new Abstract: Identity-conditioned diffusion models enable high-quality and identity-consistent face generation, but they also raise severe privacy concerns, as model

Planning in the LLM Era: Building for Reliability and Efficiency

ResearchDGX agent

arXiv:2605.21902v1 Announce Type: cross Abstract: Growing attention to intelligent agents has put a spotlight on one of their central capabilities: planning. Early attempts to leverage large language

PointLLM-R: Enhancing 3D Point Cloud Reasoning via Chain-of-Thought

ApplicationsDGX agent

arXiv:2605.22013v1 Announce Type: new Abstract: Understanding 3D point clouds through language remains a fundamental challenge in computer graphics and visual computing, due to the irregular structure

Polite on the Surface, Wrong in Practice: A Curated Dataset for Fixing Honorific Failures in Multilingual Bangla Generation

Model ReleasesDGX agent

arXiv:2605.22487v1 Announce Type: new Abstract: Recent advances in Multilingual Large Language Models (MLLMs) have significantly enhanced cross-lingual conversational capabilities, yet modeling cultur

PolycubeNet: A Dual-latent Diffusion Model for Polycube-Based Hexahedral Mesh Generation

ResearchDGX agent

arXiv:2605.20274v1 Announce Type: cross Abstract: Hexahedral meshes are widely used in simulation pipelines, yet automatic generation remains challenging for complex CAD geometries. Polycube-based hex

Pre-VLA: Preemptive Runtime Verification for Reliable Vision-Language-Action and World-Model Rollouts

Model ReleasesDGX agent

arXiv:2605.22446v1 Announce Type: new Abstract: While large vision-language-action (VLA) models and generative world models (WM) have advanced long-horizon embodied intelligence, their practical deplo

PrivacyAkinator: Articulating Key Privacy Design Decisions by Answering LLM-Generated Multiple-choice Questions

ApplicationsDGX agent

arXiv:2605.20206v1 Announce Type: cross Abstract: NIST's Privacy Risk Assessment Methodology (PRAM) provides a structured framework for privacy experts to assess privacy risks. However, its complexity

Probabilistic Attribution For Large Language Models

ResearchDGX agent

arXiv:2605.21726v1 Announce Type: new Abstract: The generative nature of Large Language Models (LLMs) is reflected in the conditional probabilities they compute to sample each response token given the

ProcBench: Evaluating Process-Level Defects and Control Preservation in LLM Coding Agents

Model ReleasesDGX agent

arXiv:2605.20251v2 Announce Type: cross Abstract: Existing benchmarks for LLM coding agents primarily evaluate final outcomes. While useful for measuring overall capability, these metrics provide limi

PromptNCE: Pointwise Mutual Information Predictions Using Only LLMs and Contrastive Estimation Prompts

Model ReleasesDGX agent

arXiv:2605.21776v1 Announce Type: new Abstract: Estimating mutual information from text usually requires training a task-specific critic, which limits its use in low-data settings. We ask whether larg

Proportional Selection in Networks

ResearchDGX agent

arXiv:2502.03545v2 Announce Type: replace-cross Abstract: We address the problem of selecting k representative nodes from a network, aiming to achieve two objectives: identifying the most influential

Psy-Chronicle:A Structured Pipeline for Synthesizing Long-Horizon Campus Psychological Counseling Dialogues

AgentsDGX agent

arXiv:2605.22140v1 Announce Type: new Abstract: In recent years, large language models have shown substantial potential in psychological support tasks. However, existing psychological counseling data

Putnam 2025 Problems in Rocq using Opus 4.6 and Rocq-MCP

Model ReleasesDGX agent

arXiv:2603.20405v2 Announce Type: replace-cross Abstract: We report on an experiment in which Claude Opus~4.6, equipped with a suite of Model Context Protocol (MCP) tools for the Rocq proof assistant,

Quality and Security Signals in AI-Generated Python Refactoring Pull Requests

AgentsDGX agent

arXiv:2605.21453v1 Announce Type: cross Abstract: As AI agents increasingly contribute to code development and maintenance, there is still limited empirical evidence on the quality and risk characteri

Quantifying Full-Body Immersion

ResearchDGX agent

arXiv:2605.22521v1 Announce Type: new Abstract: Humanity is at the forefront of yet another digital revolution, where the lines between real and virtual worlds are dissolving, reshaping how we perceiv

Quantizing Whisper-small: How design choices affect ASR performance

ApplicationsDGX agent

arXiv:2511.08093v2 Announce Type: replace-cross Abstract: Large speech recognition models like Whisper-small achieve high accuracy but are difficult to deploy on edge devices due to their high computa

QuantSR+: Pushing the Limit of Quantized Image Super-Resolution Networks

SafetyDGX agent

arXiv:2605.22351v1 Announce Type: new Abstract: Low-bit quantization is widely used to compress super-resolution (SR) models and reduce storage and computation costs for deployment on resource-limited

RAGCap-Bench: Benchmarking Capabilities of LLMs in Agentic Retrieval Augmented Generation Systems

Model ReleasesDGX agent

arXiv:2510.13910v2 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) mitigates key limitations of Large Language Models (LLMs)-such as factual errors, outdated knowledge, and hallu

RankJudge: A Multi-Turn LLM-as-a-Judge Synthetic Benchmark Generator

Model ReleasesDGX agent

arXiv:2605.21748v1 Announce Type: new Abstract: As interactive LLM-based applications are created and refined, model developers need to evaluate the quality of generated text along many possible axes.

Ratchet: A Minimal Hygiene Recipe for Self-Evolving LLM Agents

Model ReleasesDGX agent

arXiv:2605.22148v1 Announce Type: cross Abstract: Self-evolving skill libraries, pioneered by Voyager, let frozen LLM agents accumulate reusable knowledge without weight updates, yet recent evaluation

REACH: Hand Pose Estimation from Room Corners

ResearchDGX agent

arXiv:2605.22231v1 Announce Type: new Abstract: We introduce a novel 3D hand pose estimator that can accurately recover the shape and pose of people's hands in a room from afar, typically from fixed c

Real-Time Auto-Optimization in Unknown Environments via Structure-Exploiting Dual Control for Exploration and Exploitation

ResearchDGX agent

arXiv:2605.22431v1 Announce Type: new Abstract: This paper develops a fast numerical dual control for exploration and exploitation (DCEE) method to address auto-optimization problems in unknown enviro

RealUserSim: Bridging the Reality Gap in Agent Benchmarking via Grounded User Simulation

Model ReleasesDGX agent

arXiv:2605.20204v1 Announce Type: cross Abstract: LLM-based user simulation is the primary mechanism for end-to-end agent evaluation, yet simulated users are poor proxies for real humans: unconstraine

Reducing Political Manipulation with Consistency Training

SafetyDGX agent

arXiv:2605.22771v1 Announce Type: new Abstract: Large language models (LLMs) exhibit systematic political bias across a variety of sensitive contexts. We find that LLMs handle counterpart topics from

Reflecti-Mate: A Conversational Agent for Adaptive Decision-Making Support Through System 1 and System 2 Thinking

AgentsDGX agent

arXiv:2605.22509v1 Announce Type: cross Abstract: Making high-stakes personal decisions involves cognitive, emotional, and intuitive processes, and individuals differ in how they allocate attention ac

Reflective Prompt Tuning through Language Model Function-Calling

Model ReleasesDGX agent

arXiv:2605.21781v1 Announce Type: new Abstract: Large language models (LLMs) have become increasingly capable of following instructions and complex reasoning, making prompting a flexible interface for

Reinforcing VLAs in Task-Agnostic World Models

SafetyDGX agent

arXiv:2605.12334v2 Announce Type: replace Abstract: Post-training Vision-Language-Action (VLA) models via reinforcement learning (RL) in learned world models has emerged as an effective strategy to ad

← Previous
1…621622623624625…1034
Next →