AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

research

GridTimelineEvolution
19,012 results
17 Apr 2026

Style Amnesia: Investigating Speaking Style Degradation and Mitigation in Multi-Turn Spoken Language Models

ResearchDGX agent

arXiv:2512.23578v3 Announce Type: replace Abstract: In this paper, we show that when spoken language models (SLMs) are instructed to speak in a specific speaking style at the beginning of a multi-turn

Survey of Deep Learning and Physics-Based Approaches in Computational Wave Imaging

ResearchDGX agent

arXiv:2410.08329v3 Announce Type: replace Abstract: Computational wave imaging (CWI) extracts hidden structure and physical properties of a volume of material by analyzing wave signals that traverse t

The Acoustic Camouflage Phenomenon: Re-evaluating Speech Features for Financial Risk Prediction

ResearchDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.14619v1 Announce Type: cross Abstract: In computational paralinguistics, detecting cognitive load and deception from speech signals is a heavily researched domain. Recent efforts have attem

The case for fixing everything

ResearchDGX agent

The handsome new book Maintenance: Of Everything, Part One, by the tech industry legend Stewart Brand, promises to be the first in a series offering “a comprehensive overview of the civilizational imp

The Code Whisperer: LLM and Graph-Based AI for Smell and Vulnerability Resolution

ResearchDGX agent

arXiv:2604.13114v1 Announce Type: cross Abstract: Code smells and software vulnerabilities both increase maintenance cost, yet they are often handled by separate tools that miss structural context and

The Cognitive Circuit Breaker: A Systems Engineering Framework for Intrinsic AI Reliability

ResearchDGX agent

arXiv:2604.13417v1 Announce Type: cross Abstract: As Large Language Models (LLMs) are increasingly deployed in mission-critical software systems, detecting hallucinations and ``faked truthfulness'' ha

The Download: bad news for inner Neanderthals, and AI warfare’s human illusion

ResearchDGX agent

This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. The problem with thinking you’re part Neanderthal There’s a th

Threshold Differential Attention for Sink-Free, Ultra-Sparse, and Non-Dispersive Language Modeling

ResearchDGX agent

arXiv:2601.12145v2 Announce Type: replace Abstract: Softmax attention struggles with long contexts due to structural limitations: the strict sum-to-one constraint forces attention sinks on irrelevant

Tight Bounds for Learning Polyhedra with a Margin

ResearchDGX agent

arXiv:2604.14614v1 Announce Type: cross Abstract: We give an algorithm for PAC learning intersections of k halfspaces with a rho margin to within error arepsilon that runs in time extsf{poly}(k, areps

TokenFormer: Unify the Multi-Field and Sequential Recommendation Worlds

ResearchDGX agent

arXiv:2604.13737v1 Announce Type: cross Abstract: Recommender systems have historically developed along two largely independent paradigms: feature interaction models for modeling correlations among mu

TokenGS: Decoupling 3D Gaussian Prediction from Pixels with Learnable Tokens

ResearchDGX agent

arXiv:2604.15239v1 Announce Type: new Abstract: In this work, we revisit several key design choices of modern Transformer-based approaches for feed-forward 3D Gaussian Splatting (3DGS) prediction. We

TokenLight: Precise Lighting Control in Images using Attribute Tokens

ResearchDGX agent

arXiv:2604.15310v1 Announce Type: new Abstract: This paper presents a method for image relighting that enables precise and continuous control over multiple illumination attributes in a photograph. We

Towards Design Compositing

ResearchDGX agent

arXiv:2604.14605v1 Announce Type: new Abstract: Graphic design creation involves harmoniously assembling multimodal components such as images, text, logos, and other visual assets collected from diver

Towards Faster Language Model Inference Using Mixture-of-Experts Flow Matching

ResearchDGX agent

arXiv:2604.15009v1 Announce Type: cross Abstract: Flow matching retains the generation quality of diffusion models while enabling substantially faster inference, making it a compelling paradigm for ge

Tracking the Temporal Dynamics of News Coverage of Catastrophic and Violent Events

ResearchDGX agent

arXiv:2604.14315v1 Announce Type: new Abstract: The modern news cycle has been fundamentally reshaped by the rapid exchange of information online. As a result, media framing shifts dynamically as new

Trump is reportedly negotiating a deal which would stop Iran from producing a nuclear weapon in exchange for $20 Billion in Iranian assets b…

ResearchDGX agent

Trump is reportedly negotiating a deal which would stop Iran from producing a nuclear weapon in exchange for 20 Billion in Iranian assets being unfrozen. Note that The Iranian Nuclear agreement (JCPOA

Tug-of-War within A Decade: Conflict Resolution in Vulnerability Analysis via Teacher-Guided Retrieval-Augmented Generations

ResearchDGX agent

arXiv:2604.14172v1 Announce Type: new Abstract: Large Language Models (LLMs) are essential for analyzing and addressing vulnerabilities in cybersecurity. However, among over 200,000 vulnerabilities we

Universal hidden monotonic trend estimation with contrastive learning

ResearchDGX agent

arXiv:2210.09817v3 Announce Type: replace Abstract: In this paper, we describe a universal method for extracting the underlying monotonic trend factor from time series data. We propose an approach rel

Unsupervised feature selection using Bayesian Tucker decomposition

ResearchDGX agent

arXiv:2604.14949v1 Announce Type: cross Abstract: In this paper, we proposed Bayesian Tucker decomposition (BTuD) in which residual is supposed to obey Gaussian distribution analogous to linear regres

Unsupervised Learning of Local Updates for Maximum Independent Set in Dynamic Graphs

ResearchDGX agent

arXiv:2505.13754v3 Announce Type: replace Abstract: We present the first unsupervised learning model for Maximum-Independent-Set (MaxIS) in dynamic graphs where edges change over time. Our method comb

VisPCO: Visual Token Pruning Configuration Optimization via Budget-Aware Pareto-Frontier Learning for Vision-Language Models

ResearchDGX agent

arXiv:2604.15188v1 Announce Type: new Abstract: Visual token pruning methods effectively mitigate the quadratic computational growth caused by processing high-resolution images and video frames in vis

Waiting for departure 11 #pixelart

ResearchDGX agent

This is a pixel art illustration by David Ha depicting a scene of waiting for departure, likely showing characters or a scene at a transportation hub or travel location rendered in pixel art style. Th

WaveSFNet: A Wavelet-Based Codec and Spatial--Frequency Dual-Domain Gating Network for Spatiotemporal Prediction

ResearchDGX agent

arXiv:2603.23284v2 Announce Type: replace Abstract: Spatiotemporal predictive learning aims to forecast future frames from historical observations in an unsupervised manner, and is critical to a wide

When Does Content-Based Routing Work? Representation Requirements for Selective Attention in Hybrid Sequence Models

ResearchDGX agent

arXiv:2603.20997v2 Announce Type: replace Abstract: We identify a routing paradox in hybrid sequence models: content-based routing - deciding which tokens deserve expensive attention - requires pairwi

XMark: Reliable Multi-Bit Watermarking for LLM-Generated Texts

ResearchDGX agent

arXiv:2604.05242v2 Announce Type: replace Abstract: Multi-bit watermarking has emerged as a promising solution for embedding imperceptible binary messages into Large Language Model (LLM)-generated tex

Zero-Ablation Overstates Register Content Dependence in DINO Vision Transformers

ResearchDGX agent

arXiv:2604.14433v1 Announce Type: new Abstract: Zero-ablation -- replacing token activations with zero vectors -- is widely used to probe token function in vision transformers. Register zeroing in DIN

Zeroth-Order Optimization at the Edge of Stability

ResearchDGX agent

arXiv:2604.14669v1 Announce Type: new Abstract: Zeroth-order (ZO) methods are widely used when gradients are unavailable or prohibitively expensive, including black-box learning and memory-efficient f

16 Apr 2026

3DRealHead: Few-Shot Detailed Head Avatar

ResearchDGX agent

arXiv:2604.13171v1 Announce Type: new Abstract: The human face is central to communication. For immersive applications, the digital presence of a person should mirror the physical reality, capturing t

A 3D SAM-Based Progressive Prompting Framework for Multi-Task Segmentation of Radiotherapy-induced Normal Tissue Injuries in Limited-Data Settings

ResearchDGX agent

arXiv:2604.13367v1 Announce Type: new Abstract: Radiotherapy-induced normal tissue injury is a clinically important complication, and accurate segmentation of injury regions from medical images could

A closer look at how large language models trust humans: patterns and biases

ResearchDGX agent

arXiv:2504.15801v2 Announce Type: replace Abstract: As large language models (LLMs) and LLM-based agents increasingly interact with humans in decision-making contexts, understanding the trust dynamics

A Function-Centric Perspective on Flat and Sharp Minima

ResearchDGX agent

arXiv:2510.12451v2 Announce Type: replace-cross Abstract: Flat minima are strongly associated with improved generalisation in deep neural networks. However, this connection has proven nuanced in recen

A Multi-Stage Optimization Pipeline for Bethesda Cell Detection in Pap Smear Cytology

ResearchDGX agent

arXiv:2604.13939v1 Announce Type: new Abstract: Computer vision techniques have advanced significantly in recent years, finding diverse and impactful applications within the medical field. In this pap

A Multimodal Clinically Informed Coarse-to-Fine Framework for Longitudinal CT Registration in Proton Therapy

ResearchDGX agent

arXiv:2604.13397v1 Announce Type: new Abstract: Proton therapy offers superior organ-at-risk sparing but is highly sensitive to anatomical changes, making accurate deformable image registration (DIR)

A Resource-Efficient Hybrid CNN-LSTM network for image-based bean leaf disease classification

ResearchDGX agent

arXiv:2604.13835v1 Announce Type: new Abstract: Accurate and resource-efficient automated diagnosis is a cornerstone of modern agricultural expert systems. While Convolutional Neural Networks (CNNs) h

A short proof of near-linear convergence of adaptive gradient descent under fourth-order growth and convexity

ResearchDGX agent

arXiv:2604.13393v1 Announce Type: cross Abstract: Davis, Drusvyatskiy, and Jiang showed that gradient descent with an adaptive stepsize converges locally at a nearly-linear rate for smooth functions t

A transformable slender microrobot inspired by nematode parasites for interventional endovascular surgery

ResearchDGX agent

arXiv:2604.13513v1 Announce Type: new Abstract: Cardiovascular diseases account for around 17.9 million deaths per year globally, the treatment of which is challenging considering the confined space a

A Unified Conditional Flow for Motion Generation, Editing, and Intra-Structural Retargeting

ResearchDGX agent

arXiv:2604.13427v1 Announce Type: cross Abstract: Text-driven motion editing and intra-structural retargeting, where source and target share topology but may differ in bone lengths, are traditionally

A1: A Fully Transparent Open-Source, Adaptive and Efficient Truncated Vision-Language-Action Model

ResearchDGX agent

arXiv:2604.05672v3 Announce Type: replace Abstract: Vision-Language-Action (VLA) models have emerged as a powerful paradigm for open-world robot manipulation, but their practical deployment is often c

Adaptive Conformal Prediction for Improving Factuality of Generations by Large Language Models

ResearchDGX agent

arXiv:2604.13991v1 Announce Type: new Abstract: Large language models (LLMs) are prone to generating factually incorrect outputs. Recent work has applied conformal prediction to provide uncertainty es

Adaptive Learning via Off-Model Training and Importance Sampling for Fully Non-Markovian Optimal Stochastic Control. Complete version

ResearchDGX agent

arXiv:2604.13147v1 Announce Type: cross Abstract: This paper studies continuous-time stochastic control problems whose controlled states are fully non-Markovian and depend on unknown model parameters.

AI for Materials Science starter kit [D]

ResearchDGX agent

This r/MachineLearning discussion post serves as a community-curated beginner's resource for applying artificial intelligence and machine learning to materials science, likely compiling recommended to

AI-generated synthetic neurons speed up brain mapping

ResearchDGX agent

Google Research uses AI models to turn microscope images into accurate 3D neuron shapes , and generating synthetic neuron geometries helps AI learn to better classify neurons by their shape, speeding

AIM-CoT: Active Information-driven Multimodal Chain-of-Thought for Vision-Language Reasoning

ResearchDGX agent

arXiv:2509.25699v2 Announce Type: replace Abstract: Interleaved-Modal Chain-of-Thought (I-MCoT) advances vision-language reasoning, such as Visual Question Answering (VQA). This paradigm integrates sp

Any3DAvatar: Fast and High-Quality Full-Head 3D Avatar Reconstruction from Single Portrait Image

ResearchDGX agent

arXiv:2604.13856v1 Announce Type: new Abstract: Reconstructing a complete 3D head from a single portrait remains challenging because existing methods still face a sharp quality-speed trade-off: high-f

Artificial intelligence application in lymphoma diagnosis with Vision Transformer using weakly supervised training

ResearchDGX agent

arXiv:2604.13795v1 Announce Type: new Abstract: Vision transformers (ViT) have been shown to allow for more flexible feature detection and can outperform convolutional neural network (CNN) when pre-tr

Automated co-design of high-performance thermodynamic cycles via graph-based hierarchical reinforcement learning

ResearchDGX agent

arXiv:2604.13133v1 Announce Type: new Abstract: Thermodynamic cycles are pivotal in determining the efficacy of energy conversion systems. Traditional design methodologies, which rely on expert knowle

Better and Worse with Scale: How Contextual Entrainment Diverges with Model Size

ResearchDGX agent

arXiv:2604.13275v1 Announce Type: new Abstract: Larger language models become simultaneously better and worse at handling contextual information -- better at ignoring false claims, worse at ignoring i

Binomial Gradient-Based Meta-Learning for Enhanced Meta-Gradient Estimation

ResearchDGX agent

arXiv:2604.13263v1 Announce Type: new Abstract: Meta-learning offers a principled framework leveraging task-invariant priors from related tasks, with which task-specific models can be fine-tuned on do

BOAT: Navigating the Sea of In Silico Predictors for Antibody Design via Multi-Objective Bayesian Optimization

ResearchDGX agent

arXiv:2604.13980v1 Announce Type: new Abstract: Antibody lead optimization is inherently a multi-objective challenge in drug discovery. Achieving a balance between different drug-like properties is cr

Bridging Compositional and Distributional Semantics: A Survey on Latent Semantic Geometry via AutoEncoder

ResearchDGX agent

arXiv:2506.20083v4 Announce Type: replace Abstract: Integrating compositional and symbolic properties into current distributional semantic spaces can enhance the interpretability, controllability, com

C-voting: Confidence-Based Test-Time Voting without Explicit Energy Functions

ResearchDGX agent

arXiv:2604.13521v1 Announce Type: new Abstract: Neural network models with latent recurrent processing, where identical layers are recursively applied to the latent state, have gained attention as pro

C2: Scalable Rubric-Augmented Reward Modeling from Binary Preferences

ResearchDGX agent

arXiv:2604.13618v1 Announce Type: new Abstract: Rubric-augmented verification guides reward models with explicit evaluation criteria, yielding more reliable judgments than single-model verification. H

Calibrated Speculative Decoding: Frequency-Guided Candidate Selection for Efficient Inference

ResearchDGX agent

arXiv:2604.13634v1 Announce Type: new Abstract: Speculative decoding accelerates autoregressive generation by letting draft tokens bypass full verification, but conventional frameworks suffer from fre

Can Cross-Layer Transcoders Replace Vision Transformer Activations? An Interpretable Perspective on Vision

ResearchDGX agent

arXiv:2604.13304v1 Announce Type: new Abstract: Understanding the internal activations of Vision Transformers (ViTs) is critical for building interpretable and trustworthy models. While Sparse Autoenc

Caption First, VQA Second: Knowledge Density, Not Task Format, Drives Multimodal Scaling

ResearchDGX agent

arXiv:2604.13054v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have achieved rapid progress, yet their scaling behavior remains less clearly characterized and often less pred

Causal Drawbridges: Characterizing Gradient Blocking of Syntactic Islands in Transformer LMs

ResearchDGX agent

arXiv:2604.13950v1 Announce Type: new Abstract: We show how causal interventions in Transformer models provide insights into English syntax by focusing on a long-standing challenge for syntactic theor

ClipGStream: Clip-Stream Gaussian Splatting for Any Length and Any Motion Multi-View Dynamic Scene Reconstruction

ResearchDGX agent

arXiv:2604.13746v1 Announce Type: new Abstract: Dynamic 3D scene reconstruction is essential for immersive media such as VR, MR, and XR, yet remains challenging for long multi-view sequences with larg

Complex Interpolation of Matrices with an application to Multi-Manifold Learning

ResearchDGX agent

arXiv:2604.14118v1 Announce Type: new Abstract: Given two symmetric positive-definite matrices A, B in R^{n imes n}, we study the spectral properties of the interpolation A^{1-x} B^x for 0 leq x leq 1

Computational framework for multistep metabolic pathway design

ResearchDGX agent

arXiv:2604.13471v1 Announce Type: new Abstract: In silico tools are important for generating novel hypotheses and exploring alternatives in de novo metabolic pathway design. However, while many comput

Concrete Jungle: Towards Concreteness Paved Contrastive Negative Mining for Compositional Understanding

ResearchDGX agent

arXiv:2604.13313v1 Announce Type: new Abstract: Vision-Language Models demonstrate remarkable capabilities but often struggle with compositional reasoning, exhibiting vulnerabilities regarding word or

← Previous
1…291292293294295…317
Next →