AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlog
85,188Total entries
1Added by human
85,187Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
26,082 results
Research

Automatic Labelling of Speech Translation Errors

DGX agent

arXiv:2606.06047v1 Announce Type: new Abstract: Errors in speech translations reduce trustworthiness of Speech Translation (ST) systems and can have serious consequences. Yet currently there is no est

researcharxiv-cs-cl
5 Jun 2026
Research

Beyond Absolute Scores: Relative Edit-induced Difference for Generalizable Image Aesthetic Assessment

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2606.05778v1 Announce Type: new Abstract: Traditional Image Aesthetic Assessment (IAA) methods mainly rely on regressing absolute Mean Opinion Scores (MOS). However, such a paradigm overlooks th

researcharxiv-cs-cv
5 Jun 2026
Research

BrainExplore: Large-Scale Discovery of Interpretable Visual Representations in the Human Brain

DGX agent

arXiv:2512.08560v3 Announce Type: replace Abstract: Understanding how the human brain represents visual concepts, and in which brain regions these representations are encoded, remains a long-standing

researcharxiv-cs-cv
5 Jun 2026
Research

BRepCLIP: Contrastive Multimodal Pretraining on BRep Primitives for CAD Understanding

DGX agent

arXiv:2606.05515v1 Announce Type: new Abstract: Learning representations of CAD models is a largely open problem. While 3D representation learning has flourished around point clouds and meshes, the na

researcharxiv-cs-cv
5 Jun 2026
Research

CaliDist: Calibrating Large Language Models via Behavioral Robustness to Distraction

DGX agent

arXiv:2606.05799v1 Announce Type: cross Abstract: Existing calibration methods for Large Language Models (LLMs) often overlook a critical dimension of trustworthiness: a model's {em behavioral robustn

researcharxiv-cs-cl
5 Jun 2026
Research

Can We Predict The Human Preference For Text-to-Image Content Prior To Generation And Is It Even Useful To Do So?

DGX agent

arXiv:2606.05478v1 Announce Type: new Abstract: Diffusion Models (DM) have revolutionized text-driven generation by enabling the synthesis of high-quality, photorealistic visual content from user prom

researcharxiv-cs-cv
5 Jun 2026
Research

ChartAttack: Testing the Vulnerability of LLMs to Malicious Prompting in Chart Generation

DGX agent

arXiv:2601.12983v3 Announce Type: replace Abstract: Multimodal large language models (MLLMs) are increasingly used to automate chart generation from data tables, improving analysis and reporting effic

researcharxiv-cs-cl
5 Jun 2026
Research

ColBERTSaR: Sparsified ColBERT Index via Product Quantization

DGX agent

arXiv:2606.05568v1 Announce Type: cross Abstract: While ColBERT is an effective neural retrieval architecture, it requires a heavy index structure to support candidate set retrieval based on approxima

researcharxiv-cs-cl
5 Jun 2026
Research

Comparison of Deep Learning Frameworks For Rice Disease Mapping From UAV Multispectral Imaging

DGX agent

arXiv:2606.06359v1 Announce Type: new Abstract: In this study, UAV multispectral imagery is used to segment the severity of bacterial leaf blight (BLB) in rice using convolutional neural networks (CNN

researcharxiv-cs-cv
5 Jun 2026
Research

Computation-Aware Event-to-Frame Reconstruction via Selective Attention

DGX agent

arXiv:2606.06142v1 Announce Type: new Abstract: Event-to-frame (E2F) reconstruction bridges asynchronous event streams with frame-based vision pipelines, but existing methods often face a trade-off be

researcharxiv-cs-cv
5 Jun 2026
Research

Congrats to Reardon and team on @flourishailabs. If they can get AI sample efficiency and energy consumption to human levels, thats going to…

DGX agent

Congrats to Reardon and team on @flourishailabs. If they can get AI sample efficiency and energy consumption to human levels, thats going to change so many things in the world! And we are live! https:

researchsoumith-chintala--x
5 Jun 2026
Research

Correcting Prompt Dependence in LLM Benchmarks: A Bayesian Hierarchical Model with Embedding-Space Clustering

DGX agent

arXiv:2510.05709v2 Announce Type: replace-cross Abstract: LLM benchmarking metrics often misstate performance and uncertainty as they rely on two assumptions that frequently do not hold in practice: (

researcharxiv-cs-cl
5 Jun 2026
Research

CoT-Space: A Theoretical Framework for Internal Slow-Thinking via Reinforcement Learning

DGX agent

arXiv:2509.04027v3 Announce Type: replace-cross Abstract: Test-time scaling, primarily manifested through multi-step Chain-of-Thought (CoT) reasoning via Reinforcement Learning (RL), has emerged as a

researcharxiv-cs-cl
5 Jun 2026
Research

Decomposing Factual Sycophancy in Language Models: How Size and Instruction Tuning Shape Robustness

DGX agent

arXiv:2606.06306v1 Announce Type: new Abstract: Factual sycophancy occurs when a language model abandons a correct, verifiable answer under social pressure. Because a flip occurs only when pressure to

researcharxiv-cs-cl
5 Jun 2026
Research

Deep Learning-assisted AMD Staging based on OCT and OCT Angiography

DGX agent

arXiv:2606.05379v1 Announce Type: new Abstract: To develop and evaluate deep learning models for automated grading of age-related macular degeneration (AMD) severity using optical coherence tomography

researcharxiv-cs-cv
5 Jun 2026
Research

Deep Learning-based 3D Oral Cavity Reconstruction Using 2D Intraoral Images

DGX agent

arXiv:2606.05998v1 Announce Type: new Abstract: Oral 3D modelling is one of the most essential stages in dentistry, and many different approaches, such as impression taking and intraoral scanning, are

researcharxiv-cs-cv
5 Jun 2026
Research

DiG-Plan: Mitigating Early Commitment for Tool-Graph Planning via Diffusion Guidance

DGX agent

arXiv:2606.05728v1 Announce Type: cross Abstract: Generating executable tool plans requires selecting appropriate subsets from tool libraries, a combinatorial search problem with an exponentially larg

researcharxiv-cs-cl
5 Jun 2026
Research

DRIFT: A Residual Flow Adapter for Decoding Continuous Outputs in Vision-Language Models

DGX agent

arXiv:2606.05758v1 Announce Type: new Abstract: Many modern vision-language models (VLMs) build on autoregressive decoding of discrete tokens. While text-based output interfaces enable scalable pretra

researcharxiv-cs-cv
5 Jun 2026
Research

Dynamic Thinking-Token Selection for Efficient Reasoning in Large Reasoning Models

DGX agent

arXiv:2601.18383v2 Announce Type: replace-cross Abstract: Large Reasoning Models (LRMs) excel at solving complex problems by explicitly generating a reasoning trace before deriving the final answer. H

researcharxiv-cs-cl
5 Jun 2026
Research

Efficient Computation of Distance Functions for Navigation Vector Fields in Lie Groups

DGX agent

arXiv:2606.05372v1 Announce Type: new Abstract: Vector-field-based methods are widely used for robot control and are often applied to the path-tracking problem. Some vector field approaches require re

researcharxiv-cs-ro
5 Jun 2026
Research

Emotion-Aware Image Generation from Korean Diary Text via LLM-based Prompt Translation and LoRA Fine-Tuning

DGX agent

arXiv:2606.05816v1 Announce Type: new Abstract: T2I models cannot effectively capture sentiment from various types of text, including diaries, as they primarily focus on visual object-related patterns

researcharxiv-cs-cv
5 Jun 2026
Research

ExpSpeech-Net: Multimodal Fusion of Expression and Speech for Deepfake Detection

DGX agent

arXiv:2606.05760v1 Announce Type: new Abstract: Deepfake videos are increasingly challenging the credibility of online content. Many existing detection methodology relies on complex, resource-intensiv

researcharxiv-cs-cv
5 Jun 2026
Model Releases

Faithful, Enriched, and Precise: Benchmarking Natural-Science Illustration Generation by T2I models

DGX agent

arXiv:2606.05949v1 Announce Type: new Abstract: Scientific illustrations are essential tools for communicating research findings, especially in natural science, where they visualize complex concepts a

model-releasesarxiv-cs-cv
5 Jun 2026
Research

Find an important unsolved problem you care about. Then use AI to solve it. Go deep! Talk to people. Build a community. It might take you mo…

DGX agent

Find an important unsolved problem you care about. Then use AI to solve it. Go deep! Talk to people. Build a community. It might take you months or years, but always know that AI capabilities will onl

researchdair-ai--x
5 Jun 2026
Research

FontFusion: Enhancing Generative Text in Diffusion Models with Typographic Conditioning

DGX agent

arXiv:2606.06066v1 Announce Type: new Abstract: Typography generation in diffusion models faces a persistent trade-off: enabling precise font control typically degrades text legibility, while maintain

researcharxiv-cs-cv
5 Jun 2026
Research

Framing, Judging, Steering: An Assessable Competency Model for Teach-ing Students to Reason With Generative AI

DGX agent

arXiv:2606.05983v1 Announce Type: cross Abstract: Generative AI makes answers easy and understanding hard, and uncritical use invites cognitive offloading. Schools still measure unaided performance, y

researcharxiv-cs-cl
5 Jun 2026
Research

From Scoring to Explanations: Evaluating SHAP and LLM Rationales for Rubric-based Teaching Quality Assessment

DGX agent

arXiv:2606.05180v1 Announce Type: new Abstract: Automated scoring models are increasingly used to assign rubric-based quality ratings to complex language performances, including classroom transcripts,

researcharxiv-cs-cl
5 Jun 2026
Research

Global-Local Monte Carlo Tree Search in Vision-Language Models for Text-to-3D Indoor Scene Generation

DGX agent

arXiv:2606.06002v1 Announce Type: new Abstract: Large Vision-Language Models have achieved significant reasoning performance in various tasks.However, there are few studies on text-to-3D indoor scene

researcharxiv-cs-cv
5 Jun 2026
Research

GMBFormer: An NDVI-Guided Global Memory Bank Transformer for Urban Green-Space Extraction from Ultra-High-Resolution Imagery

DGX agent

arXiv:2606.06363v1 Announce Type: new Abstract: Urban green-space extraction from ultra-high-resolution (UHR) imagery is commonly performed patch by patch, which limits semantic reuse among spatially

researcharxiv-cs-cv
5 Jun 2026
Research

GRAMformer: Any-Order Modality Interactions via Volumetric Multimodal Cross-Attention

DGX agent

arXiv:2606.06249v1 Announce Type: new Abstract: Transformer-based multimodal models rely on attention mechanisms to integrate information across heterogeneous modalities. Despite their success, existi

researcharxiv-cs-cv
5 Jun 2026
Research

HDST-GNN: Heterogeneous Dynamic Spatiotemporal Graph Neural Networks for Multi-Object Tracking in UAV Aerial Imagery

DGX agent

arXiv:2606.05587v1 Announce Type: new Abstract: Multi-object tracking (MOT) from UAV imagery presents unique challenges: altitude varies across sequences, objects are small and densely packed, and fre

researcharxiv-cs-cv
5 Jun 2026
Research

Hierarchical Mask-Enhanced Dual Reconstruction Network for Few-Shot Fine-Grained Image Classification

DGX agent

arXiv:2506.20263v2 Announce Type: replace Abstract: Few-shot fine-grained image classification (FS-FGIC) is challenging as it requires distinguishing visually similar subclasses with extremely limited

researcharxiv-cs-cv
5 Jun 2026
Research

HomeWorld: A Unified Floorplan-to-Furnished Framework for Generating Controllable, Densely Interactive Whole-Home Scenes

DGX agent

arXiv:2606.06390v1 Announce Type: new Abstract: Indoor scene generation is crucial for robot simulation and modern interior design. However, complex layouts together with scarce 3D scene data make lea

researcharxiv-cs-cv
5 Jun 2026
Research

Horse Eye Blink Detection and Classification for Equine Affective State Assessment

DGX agent

arXiv:2606.05458v1 Announce Type: new Abstract: Automated detection of equine facial action units (AUs) is a promising yet under-explored avenue for pain and affective state assessment in horses. Half

researcharxiv-cs-cv
5 Jun 2026
Research

HyperVis: Continuous Latent Visual Relational Graphs on the Lorentz Hyperboloid for Compositional Reasoning

DGX agent

arXiv:2606.06100v1 Announce Type: new Abstract: Vision-Language Models (VLMs) struggle with compositional reasoning that requires understanding inter-object relationships. A natural remedy is to injec

researcharxiv-cs-cv
5 Jun 2026
Research

Imagine Before You Predict: Interleaved Latent Visual Reasoning for Video Event Prediction

DGX agent

arXiv:2606.05769v1 Announce Type: new Abstract: Video event prediction (VEP) requires models to infer unobserved future states from partial video evidence. Existing video MLLMs usually verbalize inter

researcharxiv-cs-cv
5 Jun 2026
Research

InfoDensity: Rewarding Information-Dense Traces for Efficient Reasoning

DGX agent

arXiv:2603.17310v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) with extended reasoning capabilities often generate verbose and redundant reasoning traces, incurring unnecessary

researcharxiv-cs-cl
5 Jun 2026
Research

InfoShield: Privacy-Preserving Speech Representations for Mental Health Screening via Information-Theoretic Optimization

DGX agent

arXiv:2606.05561v1 Announce Type: new Abstract: Speech-based mental health screening offers scalable depression detection, yet clinical deployment faces a significant barrier: users' privacy concerns

researcharxiv-cs-cl
5 Jun 2026
Research

Interpreting Style Representations via Style-Eliciting Prompts

DGX agent

arXiv:2606.05716v1 Announce Type: new Abstract: Style representation learning is a powerful tool for authorship analysis and modeling writing style, yet the latent nature of learned representations ma

researcharxiv-cs-cl
5 Jun 2026
Research

IR3DE: A Linear Router for Large Language Models

DGX agent

arXiv:2606.06098v1 Announce Type: new Abstract: Foundational Large Language Models (LLMs) demonstrate proficiency on a wide range of general tasks, and achieve remarkable results on various specialize

researcharxiv-cs-cl
5 Jun 2026
Research

Knowledge Distillation for Visual Autoregressive Models

DGX agent

arXiv:2606.06078v1 Announce Type: new Abstract: Autoregressive (AR) image generation models are highly expressive but computationally intensive, motivating effective model compression. Knowledge disti

researcharxiv-cs-cv
5 Jun 2026
Research

Latent Implicit Visual Reasoning

DGX agent

arXiv:2512.21218v2 Announce Type: replace Abstract: While Large Multimodal Models (LMMs) have made significant progress, they remain largely text-centric, relying on language as their core reasoning m

researcharxiv-cs-cv
5 Jun 2026
Research

Learning Contact Representation for Leg Odometry

DGX agent

arXiv:2606.05501v1 Announce Type: new Abstract: The estimation of odometry in legged robots depends on the assumption that the velocity of the foot with respect to the world remains zero during the st

researcharxiv-cs-ro
5 Jun 2026
Research

Learning from Demonstrations over Riemannian Manifolds using Neural ODEs: An Extended Abstract

DGX agent

arXiv:2606.05422v1 Announce Type: new Abstract: Learning from demonstratins (LfD) is usually performed over Euclidean spaces, while the robot state, e.g. orientation, naturally evolves over curved spa

researcharxiv-cs-ro
5 Jun 2026
Research

Learning to Route LLMs from Implicit Cost-Performance Preferences via Meta-Learning

DGX agent

arXiv:2606.06178v1 Announce Type: cross Abstract: Large language models (LLMs) present a trade-off between performance and cost, where more powerful models incur greater expense. LLM routing aims to m

researcharxiv-cs-cl
5 Jun 2026
Research

Learning What to Forget: Improving LLM Unlearning via Learned Token-Level Importance

DGX agent

arXiv:2606.06320v1 Announce Type: cross Abstract: Machine unlearning aims to remove targeted knowledge from a trained model while preserving its general capabilities. For autoregressive language model

researcharxiv-cs-cl
5 Jun 2026
Research

LLM-Conditioned Synthesis of Pathological Gaits via Structured Gait-Language Representations

DGX agent

arXiv:2606.06048v1 Announce Type: new Abstract: Pathological gait datasets remain scarce due to privacy, recruitment, cost, and movement variability. Our work presents a multimodal LLM-guided framewor

researcharxiv-cs-cv
5 Jun 2026
Research

LLM-Enhanced Dialogue Management for Full-Duplex Spoken Dialogue Systems

DGX agent

arXiv:2502.14145v3 Announce Type: replace Abstract: Achieving full-duplex communication in spoken dialogue systems (SDS) requires real-time coordination between listening, speaking, and thinking. This

researcharxiv-cs-cl
5 Jun 2026
← Previous
1…224225226227228…544
Next →