AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
22,360 results
Research

Imagine Before You Predict: Interleaved Latent Visual Reasoning for Video Event Prediction

DGX agent

arXiv:2606.05769v1 Announce Type: new Abstract: Video event prediction (VEP) requires models to infer unobserved future states from partial video evidence. Existing video MLLMs usually verbalize inter

researcharxiv-cs-cv
5 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

InfoDensity: Rewarding Information-Dense Traces for Efficient Reasoning

DGX agent

arXiv:2603.17310v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) with extended reasoning capabilities often generate verbose and redundant reasoning traces, incurring unnecessary

researcharxiv-cs-cl
5 Jun 2026
Research

InfoShield: Privacy-Preserving Speech Representations for Mental Health Screening via Information-Theoretic Optimization

DGX agent

arXiv:2606.05561v1 Announce Type: new Abstract: Speech-based mental health screening offers scalable depression detection, yet clinical deployment faces a significant barrier: users' privacy concerns

researcharxiv-cs-cl
5 Jun 2026
Research

Interpreting Style Representations via Style-Eliciting Prompts

DGX agent

arXiv:2606.05716v1 Announce Type: new Abstract: Style representation learning is a powerful tool for authorship analysis and modeling writing style, yet the latent nature of learned representations ma

researcharxiv-cs-cl
5 Jun 2026
Research

IR3DE: A Linear Router for Large Language Models

DGX agent

arXiv:2606.06098v1 Announce Type: new Abstract: Foundational Large Language Models (LLMs) demonstrate proficiency on a wide range of general tasks, and achieve remarkable results on various specialize

researcharxiv-cs-cl
5 Jun 2026
Research

Knowledge Distillation for Visual Autoregressive Models

DGX agent

arXiv:2606.06078v1 Announce Type: new Abstract: Autoregressive (AR) image generation models are highly expressive but computationally intensive, motivating effective model compression. Knowledge disti

researcharxiv-cs-cv
5 Jun 2026
Research

Latent Implicit Visual Reasoning

DGX agent

arXiv:2512.21218v2 Announce Type: replace Abstract: While Large Multimodal Models (LMMs) have made significant progress, they remain largely text-centric, relying on language as their core reasoning m

researcharxiv-cs-cv
5 Jun 2026
Research

Learning Contact Representation for Leg Odometry

DGX agent

arXiv:2606.05501v1 Announce Type: new Abstract: The estimation of odometry in legged robots depends on the assumption that the velocity of the foot with respect to the world remains zero during the st

researcharxiv-cs-ro
5 Jun 2026
Research

Learning from Demonstrations over Riemannian Manifolds using Neural ODEs: An Extended Abstract

DGX agent

arXiv:2606.05422v1 Announce Type: new Abstract: Learning from demonstratins (LfD) is usually performed over Euclidean spaces, while the robot state, e.g. orientation, naturally evolves over curved spa

researcharxiv-cs-ro
5 Jun 2026
Research

Learning to Route LLMs from Implicit Cost-Performance Preferences via Meta-Learning

DGX agent

arXiv:2606.06178v1 Announce Type: cross Abstract: Large language models (LLMs) present a trade-off between performance and cost, where more powerful models incur greater expense. LLM routing aims to m

researcharxiv-cs-cl
5 Jun 2026
Research

Learning What to Forget: Improving LLM Unlearning via Learned Token-Level Importance

DGX agent

arXiv:2606.06320v1 Announce Type: cross Abstract: Machine unlearning aims to remove targeted knowledge from a trained model while preserving its general capabilities. For autoregressive language model

researcharxiv-cs-cl
5 Jun 2026
Research

LLM-Conditioned Synthesis of Pathological Gaits via Structured Gait-Language Representations

DGX agent

arXiv:2606.06048v1 Announce Type: new Abstract: Pathological gait datasets remain scarce due to privacy, recruitment, cost, and movement variability. Our work presents a multimodal LLM-guided framewor

researcharxiv-cs-cv
5 Jun 2026
Research

LLM-Enhanced Dialogue Management for Full-Duplex Spoken Dialogue Systems

DGX agent

arXiv:2502.14145v3 Announce Type: replace Abstract: Achieving full-duplex communication in spoken dialogue systems (SDS) requires real-time coordination between listening, speaking, and thinking. This

researcharxiv-cs-cl
5 Jun 2026
Research

LLMs Can Leak Training Data But Do They Want To? A Propensity-Aware Evaluation of Memorization in LLMs

DGX agent

arXiv:2606.06286v1 Announce Type: new Abstract: Large language models can reproduce training data, but existing memorization evaluations mostly measure whether models can be forced to do so, rather th

researcharxiv-cs-cl
5 Jun 2026
Research

Many Circuits, One Mechanism: Input Variation and Evaluation Granularity in Circuit Discovery

DGX agent

arXiv:2606.06267v1 Announce Type: new Abstract: Circuit discovery methods identify subgraphs that explain specific model behaviors, and structural differences between discovered circuits are commonly

researcharxiv-cs-cl
5 Jun 2026
Research

MASF: A Multi-Model Adaptive Selection Framework for Abstractive Text summarization

DGX agent

arXiv:2606.05494v1 Announce Type: new Abstract: Automatic text summarization has become increasingly important due to the rapid growth of digital textual information. This paper presents a Multi-Model

researcharxiv-cs-cl
5 Jun 2026
Research

Measuring the sensitivity of LLM-based structured extraction to prompt, model, and schema choices in clinical discharge summaries

DGX agent

arXiv:2606.05970v1 Announce Type: new Abstract: Large language models are increasingly used for structured extraction from clinical free-text notes, but the sensitivity of their output to upstream con

researcharxiv-cs-cl
5 Jun 2026
Research

MemoryCard: Topic-Aware Multi-Modal Clue Compression for Long-Video Question Answering

DGX agent

arXiv:2606.05917v1 Announce Type: cross Abstract: Long-video question answering remains challenging for Vision-Language Models (VLMs), as answer-relevant evidence is often sparse, transient, and tempo

researcharxiv-cs-cl
5 Jun 2026
Research

Monte Carlo Steklov Operators for Large-Scale Geometry Processing in the Wild

DGX agent

arXiv:2606.05581v1 Announce Type: cross Abstract: Intrinsic methods fill the default toolbox for geometry processing on meshes. Intrinsic operators, in particular the Laplacian, underlie methods that

researcharxiv-cs-cv
5 Jun 2026
Research

MPCoT: Reward-Guided Multi-Path Latent Reasoning for Test-Time Scalable Vision-Language-Action

DGX agent

arXiv:2606.06245v1 Announce Type: new Abstract: Vision-Language-Action (VLA) policies remain brittle in long-horizon and high-uncertainty control, where one-pass action decoding provides limited infer

researcharxiv-cs-ro
5 Jun 2026
Research

MS-DKC: A Dataset Knowledge Card Framework for Designing and Adapting Medical Image Segmentation Models

DGX agent

arXiv:2606.06103v1 Announce Type: new Abstract: Medical image segmentation is often framed as a search for stronger architectures, but this can obscure a more fundamental question: what does the datas

researcharxiv-cs-cv
5 Jun 2026
Research

Multi-Granularity Reasoning for Natural Language Inference

DGX agent

arXiv:2606.05181v1 Announce Type: new Abstract: Natural Language Inference (NLI) is a fundamental task in natural language understanding that requires determining the logical relationship between a pr

researcharxiv-cs-cl
5 Jun 2026
Research

Multi-Task Crack Foundation Model for Engineering-Reliable Crack Representation and Topology Preservation in Civil Infrastructure

DGX agent

arXiv:2606.05641v1 Announce Type: new Abstract: Reliable crack assessment requires not only accurate pixel-level masks but also connected crack geometry and confidence estimates that remain stable und

researcharxiv-cs-cv
5 Jun 2026
Research

Multi-task Learning is Not Enough: Representational Entanglement in Dual-output Second Language Speech Recognition

DGX agent

arXiv:2606.06065v1 Announce Type: new Abstract: Second-language (L2) speech recognition often requires transcriptions of pronunciations and intended meanings. Multi-task learning (MTL) is a natural ap

researcharxiv-cs-cl
5 Jun 2026
Research

Multilingual Coreference Resolution via Cycle-Consistent Machine Translation

DGX agent

arXiv:2606.05444v1 Announce Type: new Abstract: Coreference resolution is a core NLP task, having a broad range of downstream applications, e.g.~machine translation, question answering, document summa

researcharxiv-cs-cl
5 Jun 2026
Research

Multilingual Detection of Alzheimer's Disease from Speech: A Cross-Linguistic Transfer Learning Approach

DGX agent

arXiv:2606.05545v1 Announce Type: new Abstract: The development of multilingual Alzheimer's Disease Dementia (AD) detection models presents significant challenges due to the resource-intensive and tim

researcharxiv-cs-cl
5 Jun 2026
Research

Multimodal Sexism Identification and Characterization using Large Language Models and Gradient Boosting

DGX agent

arXiv:2606.05997v1 Announce Type: new Abstract: We present the AILS-NTUA submission to the EXIST 2026 Lab at CLEF, addressing multimodal sexism identification and characterization in memes (Task 2) an

researcharxiv-cs-cv
5 Jun 2026
Research

Narrative Knowledge Weaver: Narrative-Centric Retrieval-Augmented Reasoning for Long-Form Text Understanding

DGX agent

arXiv:2606.05724v1 Announce Type: new Abstract: Long-form narrative QA requires reasoning over evolving story worlds rather than isolated passages: answers may depend on earlier goals, changing charac

researcharxiv-cs-cl
5 Jun 2026
Research

Next-Generation Parallel Decoder for LPDR: Architectural Optimization and Class-Balanced GAN-Augmentation

DGX agent

arXiv:2606.05785v1 Announce Type: new Abstract: Real-Time License Plate Detection and Recognition (LPDR) forms the backbone of modern smart cities. Although the YOLOV5-PDLPR model substantially improv

researcharxiv-cs-cv
5 Jun 2026
Research

ORACLE-CT: Anatomy-Aware Support Pooling for CT Classification

DGX agent

arXiv:2606.05460v1 Announce Type: new Abstract: Abdominal CT disease classification is challenging because each scan is a large 3D volume with many possible findings, while diagnostic evidence is ofte

researcharxiv-cs-cv
5 Jun 2026
Research

PAR3D: A Unified 3D-MLLM with Part-Aware Representation for Scene Understanding

DGX agent

arXiv:2606.06485v1 Announce Type: new Abstract: Recent advances in 3D multimodal large language models (3D-MLLMs) have enabled unified solutions for 3D scene understanding tasks, including visual ques

researcharxiv-cs-cv
5 Jun 2026
Research

Parallel Jacobi Decoding for Fast Autoregressive Image Generation

DGX agent

arXiv:2606.05703v1 Announce Type: new Abstract: Autoregressive (AR) models have demonstrated remarkable performance in generating high-fidelity images. However, their inherently sequential next-token

researcharxiv-cs-cv
5 Jun 2026
Research

PHUMA: Physically Reliable Humanoid Locomotion Dataset

DGX agent

arXiv:2510.26236v2 Announce Type: replace Abstract: Motion imitation is a promising approach for humanoid locomotion, enabling agents to acquire humanlike behaviors. Existing methods typically rely on

researcharxiv-cs-ro
5 Jun 2026
Research

Physics in 2-Steps: Locking Motion Priors Before Visual Refinement Erases Them

DGX agent

arXiv:2606.06361v1 Announce Type: new Abstract: Image-to-Video diffusion models leverage input images to generate visually stunning content, yet frequently produce motion that violates physical laws.

researcharxiv-cs-cv
5 Jun 2026
Research

Predictable Scaling Laws of Optimal Hyperparameters for LLM Continued Pre-training

DGX agent

arXiv:2606.05610v1 Announce Type: new Abstract: The efficacy of continued pre-training for Large Language Models (LLMs) hinges upon hyperparameter configurations, such as learning rate and batch size.

researcharxiv-cs-cl
5 Jun 2026
Research

Preserving Full 6-DOF Actuation Under Abrupt Total Rotor Failures: Passive Fault-Tolerant Flight Control Using a Biaxial-Tilt Hexacopter

DGX agent

arXiv:2606.05663v1 Announce Type: new Abstract: Conventional multirotors suffer from a rapid collapse of attainable wrench space (AWS) under abrupt total rotor failures, rendering full 6-DOF recovery

researcharxiv-cs-ro
5 Jun 2026
Research

RealDexUMI: A Wearable Universal Manipulation Interface for Dexterous Robot Learning

DGX agent

arXiv:2606.06033v1 Announce Type: new Abstract: Learning dexterous manipulation requires demonstrations that preserve fine hand-object interactions while remaining executable at deployment. Existing p

researcharxiv-cs-ro
5 Jun 2026
Research

Reinforcement Learning Elicits Contextual Learning of Unseen Language Translation

DGX agent

arXiv:2606.06428v1 Announce Type: new Abstract: Prior work has shown that large language models (LLMs) can translate unseen or low-resource languages by undergoing continued training or even by encodi

researcharxiv-cs-cl
5 Jun 2026
Research

ReTreVal: Reasoning Tree with Validation and Cross-Problem Memory for Large Language Models

DGX agent

arXiv:2601.02880v2 Announce Type: replace-cross Abstract: Every existing inference-time reasoning framework discards all failure context at problem boundaries, leaving a model solving problem 500 no w

researcharxiv-cs-cl
5 Jun 2026
Research

ReverseEOL: Improving Training-free Text Embeddings via Text Reversal in Decoder-only LLMs

DGX agent

arXiv:2606.05858v1 Announce Type: new Abstract: Recent advances in Large Language Models (LLMs) have opened new avenues for generating training-free text embeddings. However, the causal attention in d

researcharxiv-cs-cl
5 Jun 2026
Research

Revising Context, Shifting Simulated Stance: Auditing LLM-Based Stance Simulation in Online Discussions

DGX agent

arXiv:2606.06443v1 Announce Type: new Abstract: Large language models are increasingly used to simulate social media users and infer how individuals may respond to online discussions. However, it rema

researcharxiv-cs-cl
5 Jun 2026
Research

RQUL-UIE: Revitalizing Quality-Unstable Labels for Underwater Image Enhancement via In-Dataset Self-Supervision

DGX agent

arXiv:2606.06176v1 Announce Type: new Abstract: Underwater Image Enhancement (UIE) is essential for mitigating degradations caused by water medium. Although learning-based methods have advanced signif

researcharxiv-cs-cv
5 Jun 2026
Research

SAM-Flow: Source-Anchored Masked Flow for Training-Free Image Editing

DGX agent

arXiv:2606.06228v1 Announce Type: new Abstract: Training-free image editing has recently attracted increasing attention due to its ability to modify real images using powerful pre-trained diffusion an

researcharxiv-cs-cv
5 Jun 2026
Research

SC-MFJ: A Simple Haptic Quality Metric for Medical Image Segmentation

DGX agent

arXiv:2606.06199v1 Announce Type: new Abstract: Standard segmentation metrics such as Dice and Hausdorff distance measure geometric overlap but say nothing about whether a segmented surface is suitabl

researcharxiv-cs-cv
5 Jun 2026
Research

Self-Learning Expression Deformations for Data-Efficient Gaussian Avatars

DGX agent

arXiv:2606.05912v1 Announce Type: new Abstract: Modeling dynamic facial expressions using 3D Gaussian representations remains challenging due to their unstructured nature. Conventional Gaussian avatar

researcharxiv-cs-cv
5 Jun 2026
Research

Self-supervised Feature Disentanglement and Augmentation Network for One-class Face Anti-spoofing

DGX agent

arXiv:2503.22929v3 Announce Type: replace Abstract: Face anti-spoofing (FAS) techniques aim to enhance the security of facial identity authentication by distinguishing authentic live faces from decept

researcharxiv-cs-cv
5 Jun 2026
Research

Semi-Offline Reinforcement Learning for Optimized Text Generation

DGX agent

arXiv:2306.09712v2 Announce Type: replace-cross Abstract: In reinforcement learning (RL), there are two major settings for interacting with the environment: online and offline. Online methods explore

researcharxiv-cs-cl
5 Jun 2026
Research

SpanNorm: Reconciling Training Stability and Performance in Deep Transformers

DGX agent

arXiv:2601.22580v2 Announce Type: replace Abstract: The success of Large Language Models (LLMs) hinges on the stable training of deep Transformer architectures. A critical design choice is the placeme

researcharxiv-cs-cl
5 Jun 2026
← Previous
1…186187188189190…466
Next →