AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,376
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,170
  • Local Ai4,930
  • Model Releases23,883
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,376
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,170
  • Local Ai4,930
  • Model Releases23,883
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
Human
88,376Total entries
1Added by human
88,375Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,897 results
5 Jun 2026

ReasoningFlow: Discourse Structures for Understanding LLM Reasoning Traces

Model ReleasesDGX agent

arXiv:2606.05402v1 Announce Type: new Abstract: Large reasoning models (LRMs) produce reasoning traces with non-linear structures, such as backtracking and self-correction, that complicate the evaluat

ReCache: Learning Budget-Aware Caching Schedules for Diffusion Models via REINFORCE

SafetyDGX agent

arXiv:2606.06060v1 Announce Type: new Abstract: Modern diffusion models generate high-quality images and videos, but their iterative denoising process makes inference expensive. Feature caching accele

Recovering Physically Plausible Human-Object Interactions from Monocular Videos

SafetyDGX agent

arXiv:2606.05359v1 Announce Type: new Abstract: In this paper, we propose RePHO, a method to reconstruct physically plausible human-object interactions (HOI) from monocular videos. While existing kine

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

RedditPersona: A Modular Framework for Community-Conditioned LLM Adaptation from Reddit

Model ReleasesDGX agent

arXiv:2606.06027v1 Announce Type: cross Abstract: Community-conditioned language model adaptation requires choices about data collection, community definition, and evaluation that are currently made i

Reducing Hallucinations in Complex Question Answering using Simple Graph-based Retrieval-Augmented Generation (long version)

Model ReleasesDGX agent

arXiv:2606.05901v1 Announce Type: new Abstract: Large language models (LLMs) have fundamentally transformed the landscape of Natural Language Processing. Despite these advances, LLMs and LLM-based sys

Reinforcement Learning Elicits Contextual Learning of Unseen Language Translation

ResearchDGX agent

arXiv:2606.06428v1 Announce Type: new Abstract: Prior work has shown that large language models (LLMs) can translate unseen or low-resource languages by undergoing continued training or even by encodi

Representing Research Attention as Contextually Structured Flows

Model ReleasesDGX agent

arXiv:2606.05895v1 Announce Type: new Abstract: Research attention is widely used as an indicator of visibility, influence, and societal uptake, yet it is typically represented as aggregated counts th

ReSAGE-PAR: Representational Similarity Assessment for Generative Expansion in Pedestrian Attribute Recognition

SafetyDGX agent

arXiv:2606.06020v1 Announce Type: new Abstract: To address the limited diversity and data scarcity in Pedestrian Attribute Recognition (PAR), we explore image synthesis using diffusion models guided b

Resonant Minds: Closed-Loop Social Avatars with Theory of Mind

AgentsDGX agent

arXiv:2606.05896v1 Announce Type: new Abstract: Creating lifelike digital humans with genuine social intelligence requires unifying cognitive reasoning and multimodal generation within a coherent fram

Rethinking LoRA Memory Through the Lens of KV Cache Compression

Model ReleasesDGX agent

arXiv:2606.05698v1 Announce Type: new Abstract: Parametric retrieval augmentation encodes document information into lightweight, document-specific modules such as LoRA adapters, reducing the need to i

ReTreVal: Reasoning Tree with Validation and Cross-Problem Memory for Large Language Models

ResearchDGX agent

arXiv:2601.02880v2 Announce Type: replace-cross Abstract: Every existing inference-time reasoning framework discards all failure context at problem boundaries, leaving a model solving problem 500 no w

Retrospective Harness Optimization: Improving LLM Agents via Self-Preference over Trajectory Rollouts

AgentsDGX agent

arXiv:2606.05922v1 Announce Type: cross Abstract: AI agents rely on a harness of skills, tools, and workflows to solve complex problems. Continually improving this harness is essential for adapting to

ReverseEOL: Improving Training-free Text Embeddings via Text Reversal in Decoder-only LLMs

ResearchDGX agent

arXiv:2606.05858v1 Announce Type: new Abstract: Recent advances in Large Language Models (LLMs) have opened new avenues for generating training-free text embeddings. However, the causal attention in d

Revising Context, Shifting Simulated Stance: Auditing LLM-Based Stance Simulation in Online Discussions

ResearchDGX agent

arXiv:2606.06443v1 Announce Type: new Abstract: Large language models are increasingly used to simulate social media users and infer how individuals may respond to online discussions. However, it rema

Revisiting Lexicon Evaluation in Unsupervised Word Discovery

SafetyDGX agent

arXiv:2606.06183v1 Announce Type: cross Abstract: Building a lexicon from discovered word-like units is a central goal in zero-resource speech processing. But do our evaluations provide a trustworthy

RhymeFlow: Training-Free Acceleration for Video Generation with Asynchronous Denoising Flow Scheduling

TutorialsDGX agent

arXiv:2606.06309v1 Announce Type: new Abstract: Video generation models based on Diffusion Transformers (DiTs) have achieved remarkable performance in video synthesis, yet they suffer from high infere

RiskFlow: Fast and Faithful Safety-Critical Traffic Scenario Generation

SafetyDGX agent

arXiv:2606.06423v1 Announce Type: new Abstract: Safety-critical traffic scenario generation is essential for evaluating autonomous driving systems under rare but high-risk interactions. Existing diffu

Robust Scene Transfer for PointGoal Navigation via Privileged Sensor Guided Contrastive Learning

SafetyDGX agent

arXiv:2606.05506v1 Announce Type: new Abstract: We propose a sensor-guided adaptive contrastive learning framework for visual representation learning in PointGoal navigation. During training, privileg

RoCA: Robust Cross-Domain End-to-End Autonomous Driving

AgentsDGX agent

arXiv:2506.10145v3 Announce Type: replace Abstract: End-to-end (E2E) autonomous driving has recently emerged as a new paradigm, offering significant potential. However, few studies have looked into th

RQUL-UIE: Revitalizing Quality-Unstable Labels for Underwater Image Enhancement via In-Dataset Self-Supervision

ResearchDGX agent

arXiv:2606.06176v1 Announce Type: new Abstract: Underwater Image Enhancement (UIE) is essential for mitigating degradations caused by water medium. Although learning-based methods have advanced signif

Safe Embodied AI for Long-horizon Tasks: A Cross-layer Analysis of Robotic Manipulation

Model ReleasesDGX agent

arXiv:2606.05660v1 Announce Type: new Abstract: Embodied AI systems are increasingly expected to reason and act over extended horizons in physical environments. This growing capability brings safety t

SAM-Flow: Source-Anchored Masked Flow for Training-Free Image Editing

ResearchDGX agent

arXiv:2606.06228v1 Announce Type: new Abstract: Training-free image editing has recently attracted increasing attention due to its ability to modify real images using powerful pre-trained diffusion an

Sample-efficient Low-level Motion Planning for Robotic Manipulation Tasks via Zero-shot Transfer Learning

TutorialsDGX agent

arXiv:2606.06041v1 Announce Type: new Abstract: As robotic systems become more sophisticated, the growing complexity of their motion planning models and the longer training times pose substantial chal

SC-MFJ: A Simple Haptic Quality Metric for Medical Image Segmentation

ResearchDGX agent

arXiv:2606.06199v1 Announce Type: new Abstract: Standard segmentation metrics such as Dice and Hausdorff distance measure geometric overlap but say nothing about whether a segmented surface is suitabl

Scaffold, Not Vocabulary? A Controlled, Two-Tier, Pre-Registered Study of a Popperian Code-Generation Skill

Model ReleasesDGX agent

arXiv:2606.06454v1 Announce Type: cross Abstract: Large language models increasingly write, review, and judge code, and a fast-growing practice equips them with prompt 'skills' that ask the model to r

Second-order Gaussian directional derivative representations for image high-resolution corner detection

TutorialsDGX agent

arXiv:2601.08182v2 Announce Type: replace Abstract: Corner detection is widely used in various computer vision tasks, such as image matching and 3D reconstruction. Our research indicates that there ar

Seeing is Believing? Evaluating Vision-Language Model Susceptibility in Agent-to-Agent Multimodal Persuasion

Model ReleasesDGX agent

arXiv:2510.22768v2 Announce Type: replace Abstract: As autonomous agents increasingly interact, they inevitably attempt to influence one another. While prior work in text-only settings has explored th

Seeing Time: Benchmarking Chronological Reasoning and Shortcut Biases in Vision-Language Models

Model ReleasesDGX agent

arXiv:2606.05702v1 Announce Type: cross Abstract: Recent advancements in Vision-Language Models (VLMs) have significantly enhanced their ability to interpret complex visual semantics, yet their capaci

Self-Augmenting Retrieval for Diffusion Language Models

TutorialsDGX agent

arXiv:2606.06474v1 Announce Type: new Abstract: Discrete diffusion language models generate text by iteratively denoising an entire response in parallel. At each step, they predict tentative tokens fo

Self-Learning Expression Deformations for Data-Efficient Gaussian Avatars

ResearchDGX agent

arXiv:2606.05912v1 Announce Type: new Abstract: Modeling dynamic facial expressions using 3D Gaussian representations remains challenging due to their unstructured nature. Conventional Gaussian avatar

Self-supervised Feature Disentanglement and Augmentation Network for One-class Face Anti-spoofing

ResearchDGX agent

arXiv:2503.22929v3 Announce Type: replace Abstract: Face anti-spoofing (FAS) techniques aim to enhance the security of facial identity authentication by distinguishing authentic live faces from decept

Self-supervised User Profile Generation for Personalization

Model ReleasesDGX agent

arXiv:2606.05336v1 Announce Type: new Abstract: Personalizing large language models (LLMs) has become a central challenge as LLMs are deployed across recommendation, search, dialogue, and content gene

Semantic-decoupled Spatial Partition Guided Point-supervised Oriented Object Detection

HardwareDGX agent

arXiv:2506.10601v2 Announce Type: replace Abstract: Given its ability to reduce annotation costs, weakly supervised learning based on single-point annotations has emerged as a research focus in orient

Semi-Offline Reinforcement Learning for Optimized Text Generation

ResearchDGX agent

arXiv:2306.09712v2 Announce Type: replace-cross Abstract: In reinforcement learning (RL), there are two major settings for interacting with the environment: online and offline. Online methods explore

ShotCrop^3: Cropping Human-Centric Images into Cinematic Triple-Shot Compositions

Model ReleasesDGX agent

arXiv:2606.05635v1 Announce Type: new Abstract: Prior work on aesthetic composition typically produces a single aesthetically pleasing crop, overlooking the narrative value of composing multiple shots

SkillComposer: Learning to Evolve Agent Skills for Specification and Generalization

AgentsDGX agent

arXiv:2606.06079v1 Announce Type: new Abstract: Agent skills, which consist of reusable strategies that guide agent reasoning and action, have shown strong potential for improving model capability at

SoCRATES: Towards Reliable Automated Evaluation of Proactive LLM Mediation across Domains and Socio-cognitive Variations

Model ReleasesDGX agent

arXiv:2606.05563v1 Announce Type: cross Abstract: Evaluating LLM mediators remains challenging, as mediation unfolds as a real-time trajectory shaped by disputants' shifting emotions, intentions, and

SpanNorm: Reconciling Training Stability and Performance in Deep Transformers

ResearchDGX agent

arXiv:2601.22580v2 Announce Type: replace Abstract: The success of Large Language Models (LLMs) hinges on the stable training of deep Transformer architectures. A critical design choice is the placeme

Stability vs. Manipulability: Evaluating Robustness Under Post-Decision Interaction in LLM Judges

Model ReleasesDGX agent

arXiv:2606.05384v1 Announce Type: cross Abstract: LLM-as-judge evaluation is widely used in benchmarking pipelines, where model outputs are compared and ranked using automated evaluators. These pipeli

Staged Factorial Screening for Budget-Constrained Micro-Pretraining

HardwareDGX agent

arXiv:2606.05186v1 Announce Type: cross Abstract: Budget-constrained micro-pretraining often requires triaging many candidate recipes on a shared accelerator before larger search budgets are spent. We

Statistical Priors for Implicit Preferences: Decoupling Skill Selection as a Local Harness in Personal Agents

Local AiDGX agent

arXiv:2606.05828v1 Announce Type: cross Abstract: As Large Language Model (LLM) capabilities advance, locally deployed personal agents relying on API-based remote models and external skills have emerg

Statistically Reliable LLM-Based Ranking Evaluation via Prediction-Powered Inference

Model ReleasesDGX agent

arXiv:2606.05308v1 Announce Type: cross Abstract: With PRECISE, we extended Prediction-Powered Inference to produce bias-corrected estimates of ranking evaluation metrics by combining a small human-la

Staying with the Uncertainty: Uncertainty-Scaffolding Strategies for Artificial Moral Advisors in LLM-to-LLM Simulated Conversations

AgentsDGX agent

arXiv:2606.05890v1 Announce Type: new Abstract: LLMs are increasingly deployed as Artificial Moral Advisors (AMA) in a variety of contexts: what kind of conversational patterns should they display? In

StoryVideoQA: Scaling Deep Video Understanding with a Large-Scale, Multi-Genre and Auto-Generated Dataset

Model ReleasesDGX agent

arXiv:2606.06338v1 Announce Type: new Abstract: Video question answering (VideoQA) aims to answer questions about given videos. While existing approaches excel on factoid VideoQA, they struggle with d

SubtleMemory: A Benchmark for Fine-Grained Relational Memory Discrimination in Long-Horizon AI Agents

Model ReleasesDGX agent

arXiv:2606.05761v1 Announce Type: cross Abstract: Persistent AI assistants, such as OpenClaw, accumulate large collections of related memories over long-term interactions. As these memories grow, they

Symb-xMIL: Symbolic Explanations for Multiple Instance Learning in Digital Pathology

SafetyDGX agent

arXiv:2606.06224v1 Announce Type: new Abstract: Explanations of multiple instance learning (MIL) models are widely used for validation and discovery in digital histopathology. Existing methods primari

Synthetic Data Generation and Vision-based Wrinkle and Keypoint Detection for Bimanual Cloth Manipulation

ApplicationsDGX agent

arXiv:2606.06292v1 Announce Type: new Abstract: Robotic manipulation of textiles remains challenging because continuous deformation and self-occlusions hinder the robust visual perception required to

T-FunS3D: Task-Driven Hierarchical Open-Vocabulary 3D Functionality Segmentation

Local AiDGX agent

arXiv:2606.05975v1 Announce Type: new Abstract: Open-vocabulary 3D functionality segmentation enables robots to localize functional object components in 3D scenes. It is a challenging task that requir

T-SAR-JEPA: Self-Supervised Temporal Anomaly Detection in SAR Amplitude Stacks via Latent Prediction

Local AiDGX agent

arXiv:2606.05700v1 Announce Type: new Abstract: We present T-SAR-JEPA, a self-supervised framework for temporal anomaly detection in SAR amplitude stacks via latent prediction. A ViT-Base/16 encoder f

TAGA: Terrain-aware Active Gaze Learning for Generalizable Agile Humanoid Locomotion

Local AiDGX agent

arXiv:2606.05880v1 Announce Type: new Abstract: Agile humanoid locomotion across diverse challenging terrain demands both wide perceptual coverage and precise local geometry understanding. Motivated b

TAM: Torque Adaptation Module for Robust Motion Transfer in Manipulation

SafetyDGX agent

arXiv:2606.06218v1 Announce Type: new Abstract: A policy tuned for one robot often behaves differently on another, whether due to the sim-to-real gap, unknown payloads, or the differing dynamics of tw

Tamaththul3D: High-Fidelity 3D Saudi Sign Language Avatars from Monocular Video

ResearchDGX agent

arXiv:2605.05367v2 Announce Type: replace Abstract: Existing 3D sign language avatar reconstruction methods are developed and evaluated exclusively on Western sign languages, and no 3D parametric anno

TARPO: Token-Wise Latent-Explicit Reasoning via Action-Routing Policy Optimization

Model ReleasesDGX agent

arXiv:2606.05859v1 Announce Type: new Abstract: Latent reasoning has emerged as a promising alternative to discrete Chain-of-Thought (CoT) in large language models (LLMs), enabling more expressive rea

Temporal Preference Concepts and their Functions in a Large Language Model

Local AiDGX agent

arXiv:2606.05194v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly being deployed to make decisions that require trading off near-term gains against long-term consequences

TempoVLA: Learning Speed-Controllable Vision-Language-Action Policies

SafetyDGX agent

arXiv:2606.06491v1 Announce Type: new Abstract: Robot manipulation alternates between low-risk transit phases that call for fast execution and high-risk contact stages that demand slow, precise motion

Ten Headache Specialists versus Artificial Intelligence for Clinical Literature Summarization: A Critical Evaluation and Comparison

Model ReleasesDGX agent

arXiv:2606.05436v1 Announce Type: cross Abstract: Summarizing the latest medical literature to guide clinical decision-making is essential for evidence-based medicine and high-quality patient care. Ye

TensorBench: Benchmarking Coding Agents on a Compiler-Based Tensor Framework

Model ReleasesDGX agent

arXiv:2606.05570v1 Announce Type: new Abstract: Repository-level coding benchmarks face a trade-off between task difficulty and evaluation reliability: tasks that challenge frontier models often invol

Texture-preserving implicit neural representation for Cone beam CT truncated reconstruction

SafetyDGX agent

arXiv:2606.06039v1 Announce Type: new Abstract: Cone-beam computed tomography (CBCT) frequently suffers from data truncation, which introduces severe artifacts and limits the effective field of view (

TextWand: A Unified Framework for Scene Text Editing

Model ReleasesDGX agent

arXiv:2606.05730v1 Announce Type: new Abstract: We propose TextWand, a general-purpose framework that unifies scene text removal, generation, and replacement into a single model. By decomposing comple

The Generator-Eraser Paradox: Community Guidelines for Responsible LLM-Assisted Dialect Resource Creation

ApplicationsDGX agent

arXiv:2606.06004v1 Announce Type: new Abstract: Dialect resources occupy a unique position at the intersection of scientific description, cultural preservation, and computational infrastructure. Large

← Previous
1…476477478479480…1049
Next →