AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

research

GridTimelineEvolution
19,194 results
8 Jun 2026

T2LM: Long-Term 3D Human Motion Generation from Multiple Sentences

ResearchDGX agent

arXiv:2406.00636v2 Announce Type: replace Abstract: In this paper, we address the challenging problem of long-term 3D human motion generation. Specifically, we aim to generate a long sequence of smoot

TA-RAG: Tone-Aware Retrieval-Augmented Generation for Peer-Support Health Communication

ResearchDGX agent

arXiv:2606.06794v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) successfully grounds large language model (LLM) outputs in trusted documents, but factual grounding alone is insuff

TabSwift: An Efficient Tabular Foundation Model with Row-Wise Attention

ResearchDGX agent

arXiv:2606.07345v1 Announce Type: new Abstract: Tabular foundation models, exemplified by TabPFN, perform prediction via in-context learning, inferring test labels directly from labeled training examp


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Telling stories, making Hanzi: AI-assisted co-creation with elderly migrants in urban China

ResearchDGX agent

arXiv:2507.01548v3 Announce Type: replace-cross Abstract: This paper explores how older migrants in urban China can record stories that everyday language and design often miss. We ran two co-creation

Terastal: Layer-Variant-based Scheduling for Real-Time Multi-DNN Workloads on Heterogeneous Accelerators

ResearchDGX agent

arXiv:2606.06818v1 Announce Type: cross Abstract: Heterogeneous DNN accelerators improve soft real-time multi-DNN execution by mapping each layer to its preferred accelerator to reduce latency. Howeve

The Download: how the World Cup ball will fly and OpenAI’s “super app”

ResearchDGX agent

This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. Why this year’s World Cup ball may not fly as far Much is new

The Dual Mechanisms of Spatial Variable Binding in Vision-Language Models

ResearchDGX agent

arXiv:2603.22278v2 Announce Type: replace Abstract: Many multimodal tasks, such as image captioning and visual question answering, require vision-language models (VLMs) to bind objects with their prop

The Effect of Training Task Diversity on In-Context Learning through the Lens of Low-Dimensional Subspaces

ResearchDGX agent

arXiv:2606.06814v1 Announce Type: cross Abstract: The transformer's emergent ability to perform in-context learning (ICL) has sparked a wide range of studies designed to understand its underlying mech

The Geometry of Last-Layer Model Stealing

ResearchDGX agent

arXiv:2606.06854v1 Announce Type: new Abstract: This paper uses geometry to explain how a machine learning model can be stolen using an already existing well-known method. The author has shown the exa

The Identity Trap in EEG Foundation Models: A Diagnostic Audit

ResearchDGX agent

arXiv:2606.06647v1 Announce Type: new Abstract: Objective. EEG foundation models (FMs) report strong accuracy on clinical resting-state EEG. However, high accuracy under subject-disjoint cross-validat

The Latent Space: Foundation, Evolution, Mechanism, Ability, and Outlook

ResearchDGX agent

arXiv:2604.02029v2 Announce Type: replace Abstract: Latent space is rapidly emerging as a native substrate for language-based models. While modern systems are still commonly understood through explici

The Lipreading Gap: Do VSR Models Perceive Visual Speech Like Human Lipreaders?

ResearchDGX agent

arXiv:2606.07435v1 Announce Type: cross Abstract: Visual speech recognition (VSR) models now surpass human lipreaders on benchmarks, but do such gains establish human-like visual speech perception? To

The Necessity of Setting Temperature in LLM-as-a-Judge

ResearchDGX agent

arXiv:2603.28304v2 Announce Type: replace Abstract: Using large language models (LLMs) as judges for evaluating model outputs has emerged as an important paradigm for automated evaluation. However, th

The point is that you should start implementing ways to encode instructions/prompts with clear goals inside automations. Nothing new but new…

ResearchDGX agent

The point is that you should start implementing ways to encode instructions/prompts with clear goals inside automations. Nothing new but newer LLMs are being trained to perform for longer duration uni

The Proxy Benders Decomposition

ResearchDGX agent

arXiv:2606.07403v1 Announce Type: cross Abstract: Benders decomposition is a fundamental framework for solving large-scale mixed-integer optimization problems with complicating variables that, when fi

The Sharp Phase Transition of Tyler's M-Estimator for Robust Subspace Recovery

ResearchDGX agent

arXiv:2606.06782v1 Announce Type: cross Abstract: Robust Subspace Recovery (RSR) aims to identify an underlying d-dimensional subspace from a dataset heavily corrupted by outliers. Complexity-theoreti

The Utility and Complexity of in- and out-of-Distribution Machine Unlearning

ResearchDGX agent

arXiv:2412.09119v3 Announce Type: replace Abstract: Machine unlearning, the process of selectively removing data from trained models, is increasingly crucial for addressing privacy concerns and knowle

Three-dimensional hydro-cluttered locomotion by an undulatory robot

ResearchDGX agent

arXiv:2606.06829v1 Announce Type: new Abstract: Aquatic robots have expanded human access to underwater environments, yet many underwater spaces contain obstacles that can disrupt open-water locomotio

TOPSIS-RAD: Ranking According to Desires

ResearchDGX agent

arXiv:2606.07253v1 Announce Type: new Abstract: Traditional TOPSIS derives its reference points -- the Positive Ideal Solution (PIS) and Negative Ideal Solution (NIS) -- from the observed alternative

Towards Efficient and Exact Forgetting Services in Pre-Trained-Model-based Continual Learning

ResearchDGX agent

arXiv:2505.12239v2 Announce Type: replace-cross Abstract: In Continual Learning (CL), using a Pre-Trained Model (PTM) as the feature extractor has become a popular practice. Accompanied by analytic cl

TraRA: Trajectory-level Recognition Aggregation for Video Text Spotting in Urban Surveillance

ResearchDGX agent

arXiv:2606.07161v1 Announce Type: new Abstract: Video Text Spotting (VTS) is essential for urban surveillance and intelligent transportation systems, enabling automated reading of street signs, vehicl

Understanding Generative Recommendation with Semantic IDs from a Model-scaling View

ResearchDGX agent

arXiv:2509.25522v3 Announce Type: replace Abstract: Recent advancements in generative models have allowed the emergence of a promising paradigm for recommender systems (RS), known as Generative Recomm

Unified Geometry-Guided ML-FTLE for Tracking Transient Chaos from Scalar Time Series

ResearchDGX agent

arXiv:2606.07385v1 Announce Type: cross Abstract: Detecting transient chaos from scalar observations without governing equations represents a fundamental challenge in nonlinear dynamics. We propose a

Universal consistency of the k-NN rule in metric spaces and Nagata dimension. III

ResearchDGX agent

arXiv:2512.17058v3 Announce Type: replace Abstract: We establish the last missing link allowing to describe those complete separable metric spaces X in which the k nearest neighbour classifier is univ

Unmixing ATR-{mu}FTIR spectroscopic images of cross-sections of historical oil paintings

ResearchDGX agent

arXiv:2603.06673v2 Announce Type: replace Abstract: Spectroscopic imaging (SI) has become central to heritage science because it enables non-invasive, spatially resolved characterisation of materials

Unregistered Spectral Image Fusion: Unmixing, Adversarial Learning, and Recoverability

ResearchDGX agent

arXiv:2603.21510v3 Announce Type: replace-cross Abstract: This paper addresses the fusion of a pair of spatially unregistered hyperspectral image (HSI) and multispectral image (MSI) covering roughly o

Unsupervised Learning Based Focal Stack Camera Depth Estimation

ResearchDGX agent

arXiv:2203.07904v3 Announce Type: replace-cross Abstract: We propose an unsupervised deep learning based method to estimate depth from focal stack camera images. On the NYU-v2 dataset, our method achi

Varifold Moment Invariants for Sustainable and Explainable Contour Feature Extraction

ResearchDGX agent

arXiv:2606.07333v1 Announce Type: new Abstract: We introduce Varifold Moments Invariants (VMI) as a unifying framework for many previously introduced Moment Invariants. These invariants are deeply rel

Video-Based Prediction of In-Flight Particle Characteristics in Atmospheric Plasma Spraying

ResearchDGX agent

arXiv:2606.07416v1 Announce Type: new Abstract: Atmospheric plasma spraying (APS) is a widely used coating process in which in-flight particle temperature and velocity strongly influence coating quali

WAV: Multi-Resolution Block Residual Routing for Deep Decoder-Only Transformers

ResearchDGX agent

arXiv:2606.06564v1 Announce Type: cross Abstract: Residual connections are central to training deep Transformers, but standard PreNorm residual streams aggregate sublayer updates with fixed unit weigh

When Better Codebooks Are Not Enough: Predictive Performance and Behavioral Reliability in LLM Political Event Coding

ResearchDGX agent

arXiv:2606.06781v1 Announce Type: new Abstract: High accuracy does not necessarily make an LLM a faithful coder. This issue matters because many social-science studies rely on expert-written codebooks

When CLIP Sees More, It Fights Back Harder: Multi-View Guided Adaptive Counterattacks for Test-Time Adversarial Robustness

ResearchDGX agent

arXiv:2606.06938v1 Announce Type: new Abstract: Vision-language models such as CLIP have achieved remarkable zero-shot recognition capabilities, yet their robustness against adversarial perturbations

When is 3D Worth It? A Resource-Performance Frontier for CNNs and Transformers in Lung CT

ResearchDGX agent

arXiv:2606.06950v1 Announce Type: cross Abstract: Three-dimensional models are widely assumed preferable for volumetric medical imaging, yet their practical value depends on whether performance gains

Where Rectified Flows Leak: Characterising Membership Signals Along the Interpolation Path

ResearchDGX agent

arXiv:2606.07271v1 Announce Type: cross Abstract: Understanding what generative models retain from training data remains challenging, with implications for copyright and privacy. Beyond verbatim repro

Whisper Hallucination Detection and Mitigation via Hidden Representation Steering and Sparse AutoEncoders

ResearchDGX agent

arXiv:2606.07473v1 Announce Type: cross Abstract: Whisper, a widely adopted ASR model, is known to suffer from hallucinations - coherent transcriptions generated for non-speech audio entirely disconne

Why this year’s World Cup ball may not fly as far

ResearchDGX agent

Much is new about this month’s upcoming FIFA World Cup tournament, which will be held in the US, Canada, and Mexico. It hosts more teams than ever before. It’s the first to occur in three different ho

Your UnEmbedding Matrix is Secretly a Feature Lens for Text Embeddings

ResearchDGX agent

arXiv:2606.07502v1 Announce Type: new Abstract: Large language models exhibit impressive zero-shot capabilities across a wide range of downstream tasks. However, they struggle to function as off-the-s

Zero-Shot Polygon Matching with Pre-trained Models for Pose Estimation and Polygon Cloud from Challenging Stereo

ResearchDGX agent

arXiv:2511.05949v2 Announce Type: replace Abstract: While stereo matching has achieved maturity for 0D point and 1D line primitives, establishing correspondences for 2D polygons remains largely unexpl

7 Jun 2026

Member of Technical Staff (RSI Lab) https://sakana.ai/careers/member-of-technical-staff-rsi-lab/ If you are a visionary builder ready to mov…

ResearchDGX agent

Member of Technical Staff (RSI Lab) https://sakana.ai/careers/member-of-technical-staff-rsi-lab/ If you are a visionary builder ready to move to Tokyo and engineer the engine of recursive discovery, w

6 Jun 2026

A Cartography of Open Collaboration in Open Source AI: Mapping Practices, Motivations, and Governance in 14 Open Large Language Model Projects

ResearchDGX agent

arXiv:2509.25397v2 Announce Type: replace-cross Abstract: The proliferation of open large language models (LLMs) is fostering a vibrant ecosystem in artificial intelligence (AI). However, the methods

A Framework for Measuring Appropriate Reliance on Set-Valued AI Advice

ResearchDGX agent

arXiv:2606.06081v1 Announce Type: new Abstract: Appropriate reliance on AI advice has become a central research theme in human-AI collaboration. Existing frameworks have focused exclusively on point p

Adapting Diffusion Language Models for Lossless Pixel-Level Image Transmission

ResearchDGX agent

arXiv:2606.06273v1 Announce Type: cross Abstract: Lossless pixel-level image transmission is a fundamental regime beyond semantic communications, because exact recovery requires both accurate symbol p

AIS-Based Vessel Trajectory Prediction Using Memory-Augmented Neural Networks

ResearchDGX agent

arXiv:2606.06311v1 Announce Type: new Abstract: Accurate vessel trajectory prediction is essential for safe and efficient maritime operations, enabling collision avoidance and supporting route optimiz

An Improved CNN-LSTM Based Intrusion Detection System for IoT Networks

ResearchDGX agent

arXiv:2606.05776v1 Announce Type: cross Abstract: With the rapid proliferation of IoT devices, security concerns have dramatically escalated and intrusion detection systems have become critical for pr

An interpretable and trustworthy AI framework for large-scale longitudinal structure-pain association studies using data from the Osteoarthritis Initiative (OAI)

ResearchDGX agent

arXiv:2606.05357v1 Announce Type: new Abstract: Purpose: To develop an interpretable and trustworthy AI framework that combines deep learning based MRI Osteoarthritis Knee Score (MOAKS) prediction wit

Answer Presence Drives RAG Rewriting Gains

ResearchDGX agent

arXiv:2606.05633v1 Announce Type: new Abstract: Retrieval-augmented QA pipelines often route retrieved passages through an LLM rewriter before a smaller reader, lifting F1 by tens of points on multi-h

Assessing the Carbon Emissions and Energy Consumption of U.S. Hyperscale Data Centers

ResearchDGX agent

arXiv:2606.05420v1 Announce Type: new Abstract: The rapid proliferation of hyperscale data centers (HDCs) in the US, mainly driven by the adoption of artificial intelligence, has raised concerns about

Benchmarks in Leipzig

ResearchDGX agent

arXiv:2606.05818v1 Announce Type: cross Abstract: Between April 1 and May 15, 2026, a group of 49 mathematicians compiled a dataset of research-level mathematics questions with known answers. Most of

Beyond Means: Topological Causal Effects under Persistent-Homology Ignorability

ResearchDGX agent

arXiv:2603.14169v2 Announce Type: replace-cross Abstract: Average treatment effects (ATE) and conditional average treatment effects (CATE) are foundational causal estimands, but they target changes in

Beyond Vector Similarity: A Structural Analysis of Graph-Augmented Retrieval for Industrial Knowledge Graphs

ResearchDGX agent

arXiv:2606.06003v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) fails systematically on queries requiring structural reasoning over interconnected entities. We compare eight retri

Beyond Waveform Robustness: Robust Feature-Vocoder Adversarial Attacks on Automatic Speech Recognition

ResearchDGX agent

arXiv:2606.05678v1 Announce Type: cross Abstract: Automatic speech recognition (ASR) systems have become widely used for multilingual speech-to-text transcription. Their robustness to adversarial atta

Bidirectional Search for Longest Paths: Case for Front-to-Front Heuristics

ResearchDGX agent

arXiv:2606.05956v1 Announce Type: new Abstract: Bidirectional heuristic search can potentially reduce search effort for problems amenable to backward search. Therein, it is well-known that front-to-fr

Boosting Brain-to-Image Decoding with TRIBE v2 Data Augmentation

ResearchDGX agent

arXiv:2606.06345v1 Announce Type: new Abstract: Brain decoding is limited by the availability of labeled neural data, and remains challenging in low-data regimes. To address this issue, we investigate

Breaking the Chain: A Causal Analysis of LLM Faithfulness to Intermediate Structures

ResearchDGX agent

arXiv:2603.16475v2 Announce Type: replace Abstract: In schema-guided reasoning (SGR) pipelines, LLMs produce explicit intermediate structures -- rubrics, checklists, or verification queries -- before

Building a Custom Drones MuJoCo Environment [P]

ResearchDGX agent

This post likely covers the process of creating a custom drone simulation environment using MuJoCo, a physics engine commonly used in machine learning research. The project involves leveraging MuJoCo'

David Sarnoff (1891-1971) rose from Russian Jewish immigrant office boy to president/chairman of RCA. He drove commercial radio, founded NBC…

ResearchDGX agent

David Sarnoff (1891-1971) rose from Russian Jewish immigrant office boy to president/chairman of RCA. He drove commercial radio, founded NBC, and heavily funded early electronic TV development. Simila

Deciphering Two Training Clocks in Grokking via Deep Linear Network Theory with Conditional ReLU Reduction

ResearchDGX agent

arXiv:2606.05863v1 Announce Type: cross Abstract: Grokking suggests that fitting the training data and learning a simple underlying rule may occur on different time scales. We formalize this phenomeno

Dimensionality Reduction for Cyberattack Classification: A Comparative Evaluation of PCA and Linear Predictive Coding

ResearchDGX agent

arXiv:2606.05584v1 Announce Type: cross Abstract: High-dimensional feature representations are widely used in machine learning-based cyberattack detection systems. However, they increase computational

F3-Tokenizer: Taming Audio Autoencoder Latents for Understanding and Generation

ResearchDGX agent

arXiv:2606.06357v1 Announce Type: cross Abstract: Continuous audio autoencoders reconstruct waveforms well but often produce latents with weak structure for understanding, while self-supervised audio

Finite Element-Based Material Learning via Automatic Differentiation: Learning constitutive neural network models from full-field deformation data

ResearchDGX agent

arXiv:2606.05199v1 Announce Type: cross Abstract: The identification of constitutive neural network models from heterogeneous full-field deformation data provides a robust alternative to traditional c

← Previous
1…130131132133134…320
Next →