AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
22,360 results
Research

UniGeo: Unifying Geometric Guidance for Camera-Controllable Image Editing via Video Models

DGX agent

arXiv:2604.17565v1 Announce Type: new Abstract: Camera-controllable image editing aims to synthesize novel views of a given scene under varying camera poses while strictly preserving cross-view geomet

researcharxiv-cs-cv
21 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

UniMesh: Unifying 3D Mesh Understanding and Generation

DGX agent

arXiv:2604.17472v1 Announce Type: new Abstract: Recent advances in 3D vision have led to specialized models for either 3D understanding (e.g., shape classification, segmentation, reconstruction) or 3D

researcharxiv-cs-cv
21 Apr 2026
Research

Universal Diffusion-Based Probabilistic Downscaling

DGX agent

arXiv:2602.11893v3 Announce Type: replace Abstract: We introduce a universal diffusion-based downscaling framework that lifts deterministic low-resolution weather forecasts into probabilistic high-res

researcharxiv-cs-lg
21 Apr 2026
Tutorials

Unraveling the Key of Machine Learning-based Android Malware Detection

DGX agent

arXiv:2402.02953v2 Announce Type: replace-cross Abstract: With the rapid advancement of machine learning (ML), ML-based Android malware detection has gained significant popularity due to its ability t

tutorialsarxiv-cs-lg
21 Apr 2026
Research

Upper Approximation Bounds for Neural Oscillators

DGX agent

arXiv:2512.01015v2 Announce Type: replace Abstract: Neural oscillators, originating from second-order ordinary differential equations (ODEs), have demonstrated strong performance in stably learning ca

researcharxiv-cs-lg
21 Apr 2026
Research

Using Perspectival Words Is Harder Than Vocabulary Words for Humans and Even More So for Multimodal Language Models

DGX agent

arXiv:2506.00065v2 Announce Type: replace Abstract: Multimodal language models (MLMs) increasingly demonstrate human-like communication, yet their use of everyday perspectival words remains poorly und

researcharxiv-cs-cl
21 Apr 2026
Research

Variational Autoencoder Domain Adaptation for Cross-System Generalization in ML-Based SOP Monitoring

DGX agent

arXiv:2604.18035v1 Announce Type: new Abstract: Machine learning (ML) models trained to detect physical-layer threats on one optical fiber system often fail catastrophically when applied to a differen

researcharxiv-cs-lg
21 Apr 2026
Research

VC-Inspector: Advancing Reference-free Evaluation of Video Captions with Factual Analysis

DGX agent

arXiv:2509.16538v3 Announce Type: replace-cross Abstract: We propose VC-Inspector, a lightweight, open-source large multimodal model (LMM) for reference-free evaluation of video captions, with a focus

researcharxiv-cs-cl
21 Apr 2026
Research

View-Consistent 3D Scene Editing via Dual-Path Structural Correspondense and Semantic Continuity

DGX agent

arXiv:2604.17801v1 Announce Type: new Abstract: Text-driven 3D scene editing has recently attracted increasing attention. Most existing methods follow a render-edit-optimize pipeline, where multi-view

researcharxiv-cs-cv
21 Apr 2026
Research

ViPS: Video-informed Pose Spaces for Auto-Rigged Meshes

DGX agent

arXiv:2604.17623v1 Announce Type: new Abstract: Kinematic rigs provide a structured interface for articulating 3D meshes, but they lack an inherent representation of the plausible manifold of joint co

researcharxiv-cs-cv
21 Apr 2026
Research

Vision Language Models are Biased

DGX agent

arXiv:2505.23941v4 Announce Type: replace-cross Abstract: Large language models (LLMs) memorize a vast amount of prior knowledge from the Internet that helps them on downstream tasks but also may noto

researcharxiv-cs-cv
21 Apr 2026
Research

ViT^3: Unlocking Test-Time Training in Vision

DGX agent

arXiv:2512.01643v2 Announce Type: replace Abstract: Test-Time Training (TTT) has recently emerged as a promising direction for efficient sequence modeling. TTT reformulates attention operation as an o

researcharxiv-cs-cv
21 Apr 2026
Research

Vocab Diet: Reshaping the Vocabulary of LLMs via Vector Arithmetic

DGX agent

arXiv:2510.17001v2 Announce Type: replace Abstract: Large language models (LLMs) often encode word-form variation (e.g., walk vs. walked) as linear directions in the embedding space. However, standard

researcharxiv-cs-cl
21 Apr 2026
Research

Wasserstein Distributionally Robust Risk-Sensitive Estimation via Conditional Value-at-Risk

DGX agent

arXiv:2604.18546v1 Announce Type: new Abstract: We propose a distributionally robust approach to risk-sensitive estimation of an unknown signal x from an observed signal y. The unknown signal and obse

researcharxiv-cs-lg
21 Apr 2026
Research

What If Consensus Lies? Selective-Complementary Reinforcement Learning at Test Time

DGX agent

arXiv:2603.19880v2 Announce Type: replace Abstract: Test-Time Reinforcement Learning (TTRL) enables Large Language Models (LLMs) to enhance reasoning capabilities on unlabeled test streams by deriving

researcharxiv-cs-lg
21 Apr 2026
Research

What makes an entity salient in discourse?

DGX agent

arXiv:2508.16464v2 Announce Type: replace Abstract: Entities in discourse vary in salience: main participants, objects and locations stay prominent, while others are quickly forgotten, raising questio

researcharxiv-cs-cl
21 Apr 2026
Research

When Background Matters: Breaking Medical Vision Language Models by Transferable Attack

DGX agent

arXiv:2604.17318v1 Announce Type: new Abstract: Vision-Language Models (VLMs) are increasingly used in clinical diagnostics, yet their robustness to adversarial attacks remains largely unexplored, pos

researcharxiv-cs-cv
21 Apr 2026
Research

When Informal Text Breaks NLI: Tokenization Failure, Distribution Shift, and Targeted Mitigations

DGX agent

arXiv:2604.16787v1 Announce Type: new Abstract: We study how informal surface forms degrade NLI accuracy in ELECTRA-small (14M) and RoBERTa-large (355M) across four transforms applied to SNLI and Mult

researcharxiv-cs-cl
21 Apr 2026
Research

When Misinformation Speaks and Converses: Rethinking Fact-Checking in Audio Platforms

DGX agent

arXiv:2604.16767v1 Announce Type: new Abstract: Audio platforms have evolved beyond entertainment. They have become central to public discourse, from podcasts and radio to WhatsApp voice notes and liv

researcharxiv-cs-cl
21 Apr 2026
Research

When More Words Say Less: Decoupling Length and Specificity in Image Description Evaluation

DGX agent

arXiv:2601.04609v2 Announce Type: replace Abstract: Vision-language models (VLMs) are increasingly used to make visual content accessible via text-based descriptions. In current systems, however, desc

researcharxiv-cs-cl
21 Apr 2026
Research

When Seeing Overrides Knowing: Disentangling Knowledge Conflicts in Vision-Language Models

DGX agent

arXiv:2507.13868v2 Announce Type: replace Abstract: Vision-language models (VLMs) increasingly combine visual and textual information to perform complex tasks. However, conflicts between their interna

researcharxiv-cs-cv
21 Apr 2026
Research

Where is the Mind? Persona Vectors and LLM Individuation

DGX agent

arXiv:2604.17031v1 Announce Type: new Abstract: The individuation problem for large language models asks which entities associated with them, if any, should be identified as minds. We approach this pr

researcharxiv-cs-cl
21 Apr 2026
Research

Who Watches the Watchmen? Humans Disagree With Translation Metrics on Unseen Domains

DGX agent

arXiv:2604.17393v1 Announce Type: new Abstract: Automatic evaluation metrics are central to the development of machine translation systems, yet their robustness under domain shift remains unclear. Mos

researcharxiv-cs-cl
21 Apr 2026
Research

Why Training-Free Token Reduction Collapses: The Inherent Instability of Pairwise Scoring Signals

DGX agent

arXiv:2604.16745v1 Announce Type: cross Abstract: Training-free token reduction methods for Vision Transformers (ToMe, ToFu, PiToMe, and MCTF) employ different scoring mechanisms, yet they share a clo

researcharxiv-cs-cv
21 Apr 2026
Research

Writing-RL: Advancing Long-form Writing via Adaptive Curriculum Reinforcement Learning

DGX agent

arXiv:2506.05760v2 Announce Type: replace Abstract: Recent advances in Large Language Models(LLMs) have enabled strong performance in long-form writing, but current training paradigms remain limited:

researcharxiv-cs-cl
21 Apr 2026
Research

x1: Learning to Think Adaptively Across Languages and Cultures

DGX agent

arXiv:2604.16917v1 Announce Type: new Abstract: Languages encode distinct abstractions and inductive priors, yet most large language models (LLMs) overlook this diversity by reasoning in a single domi

researcharxiv-cs-cl
21 Apr 2026
Research

(1D) Ordered Tokens Enable Efficient Test-Time Search

DGX agent

arXiv:2604.15453v1 Announce Type: cross Abstract: Tokenization is a key component of autoregressive (AR) generative models, converting raw data into more manageable units for modeling. Commonly, token

researcharxiv-cs-ai
20 Apr 2026
Research

A Reconfigurable Pneumatic Joint Enabling Localized Selective Stiffening and Shape Locking in Vine-Inspired Robots

DGX agent

arXiv:2604.15907v1 Announce Type: new Abstract: Vine-inspired robots achieve large workspace coverage through tip eversion, enabling safe navigation in confined and cluttered environments. However, th

researcharxiv-cs-ro
20 Apr 2026
Research

A Structure-Preserving Graph Neural Solver for Parametric Hyperbolic Conservation Laws

DGX agent

arXiv:2604.15617v1 Announce Type: cross Abstract: Hyperbolic conservation laws govern a wide range of transport-driven dynamics featuring shocks, contact discontinuities, and complex wave interactions

researcharxiv-cs-lg
20 Apr 2026
Research

A Tale of Two Learning Algorithms: Multiple Stream Random Walk and Asynchronous Gossip

DGX agent

arXiv:2504.09792v2 Announce Type: replace Abstract: Although gossip and random walk-based learning algorithms are widely known for decentralized learning, there has been limited theoretical and experi

researcharxiv-cs-lg
20 Apr 2026
Research

Acoustic and Facial Markers of Perceived Conversational Success in Spontaneous Speech

DGX agent

arXiv:2604.15322v1 Announce Type: cross Abstract: Individuals often align their speaking patterns with their interlocutors, a phenomenon linked to engagement and rapport. While well documented in task

researcharxiv-cs-cl
20 Apr 2026
Research

Adaptive Spatio-temporal Estimation on the Graph Edges via Line Graph Transformation

DGX agent

arXiv:2311.00656v4 Announce Type: replace-cross Abstract: Spatial-temporal estimation of signals on graph edges is challenging because most conventional Graph Signal Processing techniques are defined

researcharxiv-cs-lg
20 Apr 2026
Research

Advancing Intelligent Sequence Modeling: Evolution, Trade-offs, and Applications of State- Space Architectures from S4 to Mamba

DGX agent

arXiv:2503.18970v3 Announce Type: replace Abstract: Structured State Space Models (SSMs) have emerged as a transformative paradigm in sequence modeling, addressing critical limitations of Recurrent Ne

researcharxiv-cs-lg
20 Apr 2026
Research

Aligning What Vision-Language Models See and Perceive with Adaptive Information Flow

DGX agent

arXiv:2604.15809v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have demonstrated strong capability in a wide range of tasks such as visual recognition, document parsing, and visual grou

researcharxiv-cs-cv
20 Apr 2026
Research

Anthropomorphism and Trust in Human-Large Language Model interactions

DGX agent

arXiv:2604.15316v1 Announce Type: cross Abstract: With large language models (LLMs) becoming increasingly prevalent in daily life, so too has the tendency to attribute to them human-like minds and emo

researcharxiv-cs-ai
20 Apr 2026
Research

ArrayTac: A Closed-loop Piezoelectric Tactile Platform for Continuously Tunable Rendering of Shape, Stiffness, and Friction

DGX agent

arXiv:2603.13829v2 Announce Type: replace-cross Abstract: Human touch depends on the integration of shape, stiffness, and friction, yet existing tactile displays cannot render these cues together as c

researcharxiv-cs-ai
20 Apr 2026
Research

AST: Adaptive, Seamless, and Training-Free Precise Speech Editing

DGX agent

arXiv:2604.16056v1 Announce Type: cross Abstract: Text-based speech editing aims to modify specific segments while preserving speaker identity and acoustic context. Existing methods rely on task-speci

researcharxiv-cs-ai
20 Apr 2026
Research

Attention Sinks Are Provably Necessary in Softmax Transformers: Evidence from Trigger-Conditional Tasks

DGX agent

arXiv:2603.11487v5 Announce Type: replace Abstract: Transformers often display an attention sink: probability mass concentrates on a fixed, content-agnostic position. Are sinks a byproduct of the opti

researcharxiv-cs-lg
20 Apr 2026
Research

ATTNPO: Attention-Guided Process Supervision for Efficient Reasoning

DGX agent

arXiv:2602.09953v2 Announce Type: replace Abstract: Large reasoning models trained with reinforcement learning and verifiable rewards (RLVR) achieve strong performance on complex reasoning tasks, yet

researcharxiv-cs-cl
20 Apr 2026
Research

Author-in-the-Loop Response Generation and Evaluation: Integrating Author Expertise and Intent in Responses to Peer Review

DGX agent

arXiv:2602.11173v2 Announce Type: replace Abstract: Author response (rebuttal) writing is a critical stage of scientific peer review that demands substantial author effort. In practice, authors posses

researcharxiv-cs-cl
20 Apr 2026
Research

Beyond Surface Statistics: Robust Conformal Prediction for LLMs via Internal Representations

DGX agent

arXiv:2604.16217v1 Announce Type: cross Abstract: Large language models are increasingly deployed in settings where reliability matters, yet output-level uncertainty signals such as token probabilitie

researcharxiv-cs-ai
20 Apr 2026
Research

Beyond Text Prompts: Precise Concept Erasure through Text-Image Collaboration

DGX agent

arXiv:2604.15829v1 Announce Type: new Abstract: Text-to-image generative models have achieved impressive fidelity and diversity, but can inadvertently produce unsafe or undesirable content due to impl

researcharxiv-cs-cv
20 Apr 2026
Research

BioHiCL: Hierarchical Multi-Label Contrastive Learning for Biomedical Retrieval with MeSH Labels

DGX agent

arXiv:2604.15591v1 Announce Type: cross Abstract: Effective biomedical information retrieval requires modeling domain semantics and hierarchical relationships among biomedical texts. Existing biomedic

researcharxiv-cs-ai
20 Apr 2026
Research

Brain Score Tracks Shared Properties of Languages: Evidence from Many Natural Languages and Structured Sequences

DGX agent

arXiv:2604.15503v1 Announce Type: new Abstract: Recent breakthroughs in language models (LMs) using neural networks have raised the question: how similar are these models' processing to human language

researcharxiv-cs-cl
20 Apr 2026
Research

Breakout-picker: Reducing false positives in deep learning-based borehole breakout characterization from acoustic image logs

DGX agent

arXiv:2604.16011v1 Announce Type: new Abstract: Borehole breakouts are stress-induced spalling on the borehole wall, which are identifiable in acoustic image logs as paired zones with near-symmetry az

researcharxiv-cs-cv
20 Apr 2026
Research

Chain-of-Thought Degrades Visual Spatial Reasoning Capabilities of Multimodal LLMs

DGX agent

arXiv:2604.16060v1 Announce Type: cross Abstract: Multimodal Reasoning Models (MRMs) leveraging Chain-of-Thought (CoT) based thinking have revolutionized mathematical and logical problem-solving. Howe

researcharxiv-cs-ai
20 Apr 2026
Research

CIG: Measuring Conversational Information Gain in Deliberative Dialogues with Semantic Memory Dynamics

DGX agent

arXiv:2604.15647v1 Announce Type: new Abstract: Measuring the quality of public deliberation requires evaluating not only civility or argument structure, but also the informational progress of a conve

researcharxiv-cs-cl
20 Apr 2026
Research

CiPO: Counterfactual Unlearning for Large Reasoning Models through Iterative Preference Optimization

DGX agent

arXiv:2604.15847v1 Announce Type: new Abstract: Machine unlearning has gained increasing attention in recent years, as a promising technique to selectively remove unwanted privacy or copyrighted infor

researcharxiv-cs-cl
20 Apr 2026
← Previous
1…356357358359360…466
Next →