AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

research

GridTimelineEvolution
18,859 results
6 Aug 2026

Towards End-to-End Multilingual Metaphor Processing: Integrating Detection, Translation, and Evaluation

ResearchDGX agent

arXiv:2608.04260v1 Announce Type: new Abstract: Metaphorical language remains a major challenge for multilingual natural language processing because successful interpretation and translation require r

Towards Understanding Gradient Flow Dynamics of Homogeneous Neural Networks Beyond the Origin

ResearchDGX agent

arXiv:2502.15952v3 Announce Type: replace Abstract: Recent works exploring the training dynamics of homogeneous neural network weights under gradient flow with small initialization have established th

Towards Valid B-Rep Generation: Training-Free Wireframe Anomaly Detection and Repair

ResearchDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2608.04955v1 Announce Type: new Abstract: Multi-stage boundary representation (B-Rep) generation leverages intermediate wireframes to synthesize CAD models. However, geometric and topological ri

Transferable Dual-Stream Representations for Mesoscale-Preserving Sea Surface Temperature Downscaling

ResearchDGX agent

arXiv:2608.04230v1 Announce Type: cross Abstract: Deep learning models for scientific spatio-temporal downscaling often minimize reconstruction error while failing to preserve physically meaningful mu

TRCoRSurg: Temporal-Relational Co-Reasoning for Surgical Video Triplet Recognition

ResearchDGX agent

arXiv:2608.04606v1 Announce Type: new Abstract: Understanding complex surgical scenes requires recognizing multiple interdependent entities, such as instruments, actions, and targets, while maintainin

TRNet: Topography-Guided Frequency Rectification and Structure-Aware Decoding for Multimodal Paddy Rice Segmentation

ResearchDGX agent

arXiv:2608.04154v1 Announce Type: cross Abstract: Mapping paddy rice from very-high-resolution imagery in mountainous and hilly regions is difficult because terrain alters optical appearance and incre

TS2TabPFN: Time Series Classification and Extrinsic Regression through Feature Extraction and a Tabular Foundation Model

ResearchDGX agent

arXiv:2608.04174v1 Announce Type: new Abstract: Time series data are ubiquitous in practical applications, where classification (TSC) and extrinsic regression (TSER) have emerged as essential tasks fo

Uncertainty-aware Predict-Then-Optimize Framework for Equitable Post-Disaster Power Restoration

ResearchDGX agent

arXiv:2508.04780v2 Announce Type: replace-cross Abstract: The increasing frequency of extreme weather events, such as hurricanes, highlights the urgent need for efficient and equitable power system re

Understanding Fault Tolerance of Adversarially Robust Pruned Models

ResearchDGX agent

arXiv:2608.04173v1 Announce Type: new Abstract: Deep neural networks (DNNs) deployed on resource-constrained neuromorphic hardware face three concurrent challenges: the need for model compression thro

Unifying quantum measurement constructions via a relative-entropy minimum change principle

ResearchDGX agent

arXiv:2608.04055v1 Announce Type: cross Abstract: The minimum change principle provides an information-theoretic characterization of the Bayes reversal channel in classical probability theory and has

Unleashing the Potential of Vision-Language Models for Generalizable AI-Generated Image Detection

ResearchDGX agent

arXiv:2608.04935v1 Announce Type: new Abstract: Recent work has shown that a simple linear probe on frozen representations from modern vision foundation models (VFMs) can achieve state-of-the-art AIGI

Variational Bounds for Perceptron Learning from Structured Data

ResearchDGX agent

arXiv:2608.04882v1 Announce Type: new Abstract: We introduce a variational approach to a finite-temperature continuous-spin perceptron trained on a Gaussian mixture. The model allows for a broad class

Visual Anchoring in Diffusion: Multimodal Zero-Shot Skeleton Action Recognition

ResearchDGX agent

arXiv:2608.04623v1 Announce Type: new Abstract: Zero-shot Skeleton Action Recognition (ZSAR) remains ambiguous when unseen actions share similar skeleton joint dynamics but differ in objects or scene

When Absence Is Evidence: Evaluating Completeness-Sensitive Negative Reasoning in Large Language Models

ResearchDGX agent

arXiv:2608.04591v1 Announce Type: cross Abstract: Large language models (LLMs) are often asked whether something is absent from a record, list, or retrieved context. Yet non-observation licenses a neg

When does training on downscaled images yield the same gradients?

ResearchDGX agent

arXiv:2608.04448v1 Announce Type: cross Abstract: Diffusion transformers deliver strong image generation, but their training cost grows superlinearly with resolution. Recent work justifies training or

When Large Language Models Know the Table: A Framework for Assessing Data Contamination in Tabular Datasets

ResearchDGX agent

arXiv:2510.20351v3 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly exposed to data contamination, i.e., performance gains driven by prior exposure of test datasets

When Modalities Fail to Tango: Conformal Backdoor Detection in Multimodal Contrastive Learning

ResearchDGX agent

arXiv:2608.04052v1 Announce Type: cross Abstract: Backdoor attacks in multimodal contrastive learning (MCL) have garnered growing attention in recent years, as many downstream tasks critically depend

When More Becomes Less: Position-Dependent Repetition Effects in Language Models

ResearchDGX agent

arXiv:2608.04021v1 Announce Type: new Abstract: Cloze-style probes that vary how often a target token appears implicitly assume that more copies of a target affect prediction the same way regardless o

When Proxy Prediction Becomes Equation Reconstruction: Diagnostics and Residual Learning for Factor-Derived Proxy Supervision

ResearchDGX agent

arXiv:2608.04393v1 Announce Type: new Abstract: Scientific machine learning often relies on proxy targets computed from known domain factors when direct observations are limited. When those same facto

YOLO-PVC: 2D-to-3D Consolidation of Slice-wise Detections for Volumetric Liver Tumor Localization in MRI

ResearchDGX agent

arXiv:2608.04642v1 Announce Type: new Abstract: Slice-wise 2D object detectors are increasingly applied to volumetric data due to their computational efficiency and scalability, yet they often yield f

5 Aug 2026

A Direct Route to Markov Chain Convergence via Asymptotic Equivalence with the Target

ResearchDGX agent

arXiv:2608.03353v1 Announce Type: cross Abstract: For a Markov kernel T with an invariant probability measure pi, we give a self-contained proof of the Markov chain convergence theorem via a criterion

A Human-in-the-Loop Deep Learning Framework for Color Reconstruction of Lenticular Films

ResearchDGX agent

arXiv:2608.02835v1 Announce Type: new Abstract: Historical lenticular films, such as those created with the Kodacolor process, encode color information in a distinctive spatial format. This structure

A Hyperfinite Framework for Score-Based Generative Modeling

ResearchDGX agent

arXiv:2608.02799v1 Announce Type: cross Abstract: Score-based diffusion models are typically formulated using continuous-time stochastic differential equations and measure-theoretic stochastic calculu

A Low-Cost Hybrid Reservoir Computing Model for Isolated Sign Language Video Recognition

ResearchDGX agent

arXiv:2608.03444v1 Announce Type: cross Abstract: Sign language recognition (SLR) enhances communication between hearing and hearing-impaired individuals. Although deep learning (DL) has achieved prom

A machine-readable catalogue of the Tsiolkovsky papers (fond 555, Archive of the Russian Academy of Sciences), and a way to measure how well its handwriting can be read

ResearchDGX agent

arXiv:2608.03617v1 Announce Type: new Abstract: The personal archive of Konstantin Tsiolkovsky (1857-1935) is held as fond 555 of the Archive of the Russian Academy of Sciences. The archive scanned th

A Survey on Design Methodologies for Accelerating Deep Learning on Heterogeneous Architectures

ResearchDGX agent

arXiv:2311.17815v3 Announce Type: replace-cross Abstract: Given their increasing size and complexity, the need for efficient execution of deep neural networks has become increasingly pressing in the d

A Theory of Conditional Collapse under Low-Rank Weight-Space Ablations: I. The Single-Block Theory and Synthetic Validation

ResearchDGX agent

arXiv:2608.03620v1 Announce Type: cross Abstract: Activation patching and weight-space ablation both claim a component is causally responsible for a behavior, yet they act on different objects: one fo

A Unified Resolution-Conditioned Framework for Orthogonal Line-Scanning Image Fusion

ResearchDGX agent

arXiv:2608.03107v1 Announce Type: new Abstract: Laser line-scanning microscopy enables fast volumetric imaging but produces anisotropic lateral resolution. Orthogonal line scans provide complementary

Activation-Guided Neuron Intervention to Induce Alzheimer's-Related Computational Language Phenotypes in a Large Language Model

ResearchDGX agent

arXiv:2608.03067v1 Announce Type: new Abstract: Changes in spontaneous speech provide an early signal of cognitive dysfunction in Alzheimer's disease (AD) that large language models (LLMs) can detect.

Active Stiffness Control of a Supportive Continuum Robot

ResearchDGX agent

arXiv:2608.03677v1 Announce Type: new Abstract: Supportive continuum robots (SCRs) enhance the load-bearing capability of an operative continuum robot by mechanically coupling it with a supportive arm

Adaptive Modality Reliability Diagnosis and Restoration for Robust Multimodal Intent Recognition

ResearchDGX agent

arXiv:2608.03475v1 Announce Type: cross Abstract: Multimodal intent recognition combines linguistic, acoustic, and visual evidence, but individual modalities may be noisy, missing, semantically confli

Adversarial Purification by Consistency-aware Latent Space Optimization on Data Manifolds

ResearchDGX agent

arXiv:2412.08394v2 Announce Type: replace Abstract: Deep neural networks (DNNs) are vulnerable to adversarial samples crafted by adding imperceptible perturbations to clean data, potentially leading t

AI-Based Sound Effect Generation: A Narrative Review of Generative Models Across Input Modalities

ResearchDGX agent

arXiv:2608.03742v1 Announce Type: cross Abstract: Sound effects play a crucial role in conveying actions, events, and environmental cues across digital applications, often requiring a high degree of v

AI Forensics Across White-, Grey-, and Black-Box Access: A Process Model and Research Agenda for Post-Incident Investigation of AI Systems

ResearchDGX agent

arXiv:2608.03520v1 Announce Type: cross Abstract: AI systems are increasingly involved in decisions and actions that may later require investigation. When an AI related incident occurs, investigators

Ai2 expands collaboration with Hugging Face to accelerate open science

ResearchDGX agent

Ai2 is expanding its partnership with Hugging Face to give its growing portfolio of fully open models, datasets, benchmarks, and applications the storage, bandwidth, and integrations needed to reach m

AIDE: Automated Instruction via Distilled Expertise for Reference-Free Motor Skill Coaching

ResearchDGX agent

arXiv:2608.03047v1 Announce Type: new Abstract: Generating natural-language coaching feedback on motor skills can accelerate learning, yet expert coaches are scarce and expensive. Existing reference-b

AnchorKV: Anchor-Residual KV Cache Compression

ResearchDGX agent

arXiv:2608.02901v1 Announce Type: cross Abstract: The key-value (KV) cache is the primary memory bottleneck in long-context LLM inference. Existing approaches attack it from opposite ends: eviction me

ARCHead: Activation-Metric Residual Correction for Large Language Model Output Heads

ResearchDGX agent

arXiv:2608.02703v1 Announce Type: new Abstract: Weight-only quantization substantially reduces the storage of large language model (LLM) transformer blocks, but practical backends often retain the fin

Assessing speech quality metrics for evaluation of neural audio codecs under clean speech conditions

ResearchDGX agent

arXiv:2509.24457v1 Announce Type: cross Abstract: Objective speech-quality metrics are widely used to assess codec performance. However, for neural codecs, it is often unclear which metrics provide re

Assessing the Effect of Cross-Domain Mapping on Creativity in Humans and Large Language Models

ResearchDGX agent

arXiv:2603.19087v2 Announce Type: replace Abstract: Creativity is the ability to come up with novel ideas, a capacity crucial for human development and flourishing. Are large language models (LLMs) cr

Attention is Case-Sensitive

ResearchDGX agent

arXiv:2608.03711v1 Announce Type: cross Abstract: In human visual perception, uppercase lettering serves as a natural salience cue that captures attention within lowercase text. In this paper, we pres

Automatic Patient-Specific Microwave Ablation Planning Accelerated by a Physics-Guided Deep Learning Model

ResearchDGX agent

arXiv:2608.03086v1 Announce Type: cross Abstract: Microwave ablation (MWA) is a promising minimally invasive treatment for liver tumors, but its therapeutic outcome strongly depends on patient-specifi

Bayesian Data Reweighting Improves Multimodal Retrieval for Knowledge-Based Visual Question Answering

ResearchDGX agent

arXiv:2608.02907v1 Announce Type: new Abstract: Multimodal retrievers are essential for knowledge-based visual question answering, where they retrieve external evidence for image-question pairs. Howev

Behaviorally Adaptive Visual Diversion for Inclusive and Resilient Digital Assessment Delivery

ResearchDGX agent

arXiv:2608.03531v1 Announce Type: new Abstract: Institutions increasingly rely on browser lockdown, webcam monitoring, and behavioral analytics to secure high-stakes digital assessments, yet these mec

Benign interpolation and Occam's razor

ResearchDGX agent

arXiv:2608.03386v1 Announce Type: new Abstract: Contemporary deep learning methods generalize well even when they fit their training data perfectly, a phenomenon known as benign interpolation. This ph

Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation

ResearchDGX agent

arXiv:2608.02791v1 Announce Type: new Abstract: MLLM-based segmentation faces a core segmentation trilemma: high segmentation performance, preserved dialogue ability, and fast inference. Embedding-pre

Beyond Accuracy: A Multidimensional Evaluation of Statistical Reasoning in Large Language Models

ResearchDGX agent

arXiv:2608.03038v1 Announce Type: new Abstract: Statistical reasoning is multidimensional, yet evaluations of large language models (LLMs) typically emphasize response accuracy while overlooking how m

Beyond the Hivemind: Escaping LLM Homogeneity via Meta-Persona Anchoring and Sequential Temperature Scaling

ResearchDGX agent

arXiv:2608.02618v1 Announce Type: new Abstract: Recent studies have identified an ``Artificial Hivemind'' effect in Large Language Models (LLMs) causing models to converge on a narrow, homogenized con

Bi-Lipschitz Ansatz for Anti-Symmetric Functions

ResearchDGX agent

arXiv:2503.04263v2 Announce Type: replace Abstract: Motivated by applications to the simulation of quantum many-body systems by neural networks, researchers have suggested several models which are ant

Bi-semantic Chemical Embedder for Joint Representation Learning of SMILES and Natural Language

ResearchDGX agent

arXiv:2608.03855v1 Announce Type: new Abstract: Transformer models have revolutionized natural language processing (NLP), and text-based molecular representations like SMILES have successfully extende

Biconvex Optimization for Smooth Minimum-Time Trajectories around Convex Obstacles

ResearchDGX agent

arXiv:2608.02834v1 Announce Type: new Abstract: We present a biconvex approach for minimum-time motion planning around convex obstacles that is guaranteed to converge, is anytime, and supports derivat

Calibrating Semantic Uncertainty from Observable Language-Model Probabilities

ResearchDGX agent

arXiv:2607.17447v2 Announce Type: replace-cross Abstract: As generative artificial intelligence enters scientific and professional work, its uncertainty must be defined on the states that matter for i

Can Training Logs Make Model Comparisons More Precise?

ResearchDGX agent

arXiv:2608.02705v1 Announce Type: cross Abstract: Comparing stochastically trained models requires estimating both a performance difference and its uncertainty from repeated runs. We study whether tra

Caved or Convinced: Temporal Sampling Gates Claim Deference in Video Large Language Models

ResearchDGX agent

arXiv:2608.03160v1 Announce Type: cross Abstract: When asked which of two events came first, video large language models can fail in two opposite ways: cave to a false claim, or reject a true one. Pri

Character Iconicity vs. Arbitrariness: An Arabic NLP Perspective

ResearchDGX agent

arXiv:2608.02935v1 Announce Type: new Abstract: Arabic script uses 28 letters, many of which share a common base shape (rasm) and are distinguished only by dot placement. Because early Arabic manuscri

Chat Debugging: An Exploratory Study of Human-AI Collaboration to Debug Analog Circuits

ResearchDGX agent

arXiv:2608.02955v1 Announce Type: cross Abstract: This research paper describes an exploratory study on the effectiveness of Chat Debugging: troubleshooting malfunctioning analog circuits on breadboar

ChronoLens: Measuring Language Change Across Time, Languages, and Linguistic Levels

ResearchDGX agent

arXiv:2608.03507v1 Announce Type: cross Abstract: Historical language change affects morphology, syntax, semantics, and pragmatics, yet computational studies typically examine these levels with incomp

Clarity Contrast and Similarity Selection for Multi-Focus Image Fusion

ResearchDGX agent

arXiv:2608.03252v1 Announce Type: new Abstract: Multi-focus image fusion (MFIF) aims to generate an all-in-focus image from multiple images of the same scene focused at different regions. Most existin

CLIP-EBC: CLIP Can Count Accurately through Enhanced Blockwise Classification

ResearchDGX agent

arXiv:2403.09281v3 Announce Type: cross Abstract: We propose CLIP-EBC, the first fully CLIP-based model for accurate crowd density estimation. While the CLIP model has demonstrated remarkable success

Compass: Degradation-Simulated Reciprocal Learning with Lightweight Needle RWKV for Multimodal Crack Segmentation under Missing Modalities

ResearchDGX agent

arXiv:2608.03559v1 Announce Type: new Abstract: In multimodal crack segmentation for industrial facilities, the key challenge is preventing missing modalities from degrading pixel-level performance wh

← Previous
1…1617181920…315
Next →