AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

research

GridTimelineEvolution
19,016 results
28 Apr 2026

Instruction-Free Tuning of Large Vision Language Models for Medical Instruction Following

ResearchDGX agent

arXiv:2603.19482v2 Announce Type: replace Abstract: Large vision language models (LVLMs) have demonstrated impressive performance across a wide range of tasks. These capabilities largely stem from vis

Integrative neurocybernetic modeling in the era of large-scale neuroscience

ResearchDGX agent

arXiv:2604.23903v1 Announce Type: cross Abstract: Large-scale neuroscience is generating rich datasets across animals, brain areas and behavioral contexts, yet our modeling efforts remains fragmented

IntentVLM: Open-Vocabulary Intention Recognition through Forward-Inverse Modeling with Video-Language Models

ResearchDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.24002v1 Announce Type: cross Abstract: Improving the effectiveness of human-robot interaction requires social robots to accurately infer human goals through robust intention understanding.

Interoceptive machine framework: Toward interoception-inspired regulatory architectures in artificial intelligence

ResearchDGX agent

arXiv:2604.24527v1 Announce Type: new Abstract: This review proposes an integrative framework grounded on interoception and embodied AI-termed the interoceptive machine framework-that translates biolo

Interpretable Physics-Informed Load Forecasting for U.S. Grid Resilience: SHAP-Guided Ensemble Validation in Hybrid Deep Learning Under Extreme Weather

ResearchDGX agent

arXiv:2604.23500v1 Announce Type: cross Abstract: Accurate short-term electricity load forecasting is a cornerstone of U.S. grid reliability; however, prevailing deep learning models remain opaque, li

ISExplore:Informative Segment Selection for Efficient Personalized 3D Talking Face Generation

ResearchDGX agent

arXiv:2511.07940v2 Announce Type: replace Abstract: Talking Face Generation (TFG) methods based on Neural Radiance Fields (NeRF) and 3D Gaussian Splatting (3DGS) have recently achieved impressive prog

It’s remarkable. Even with small models, Hermes just wrecks OpenClaw. But it makes sense! The team is underrated when it comes to having som…

ResearchDGX agent

It’s remarkable. Even with small models, Hermes just wrecks OpenClaw. But it makes sense! The team is underrated when it comes to having some of the first cracked prompt engineers in the game. They in

@jon_barron 'World models' has a technical meaning - the transition model/dynamics model from Bellman/Kalman in the context of MDPs/ state s…

ResearchDGX agent

@jon_barron 'World models' has a technical meaning - the transition model/dynamics model from Bellman/Kalman in the context of MDPs/ state space approach to control theory ~ 1960. I gave a talk on thi

K-SENSE: A Knowledge-Guided Self-Augmented Encoder for Neuro-Semantic Evaluation of Mental Health Conditions on Social Media

ResearchDGX agent

arXiv:2604.23493v1 Announce Type: cross Abstract: Early detection of mental health conditions, particularly stress and depression, from social media text remains a challenging open problem in computat

Karpathy dropped a 200-line GPT, so I used the math to turn pandas DataFrames into searchable context windows and open sourced it (and automated my stats pipeline). [P]

ResearchDGX agent

This post describes a project where the author leveraged Andrej Karpathy's compact GPT implementation to create a tool for converting pandas DataFrames into searchable context windows, which they then

KERV: Kinematic-Rectified Speculative Decoding for Embodied VLA Models

ResearchDGX agent

arXiv:2603.01581v2 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models build a token-domain robot control paradigm, yet suffer from low speed. Speculative Decoding (SD) is an op

Keypoint-based Dynamic Object 6-DoF Pose Tracking via Event Camera

ResearchDGX agent

arXiv:2604.23387v1 Announce Type: new Abstract: Accurate 6-DoF pose estimation of objects is critical for robots to perform precise manipulation tasks. However, for dynamic object pose estimation, con

Knee-xRAI: An Explainable AI Framework for Automatic Kellgren-Lawrence Grading of Knee Osteoarthritis

ResearchDGX agent

arXiv:2604.23435v1 Announce Type: cross Abstract: Radiographic grading of knee osteoarthritis (KOA) with the Kellgren-Lawrence (KL) system is limited by inter-reader variability and the opacity of cur

Knowledge Vector of Logical Reasoning in Large Language Models

ResearchDGX agent

arXiv:2604.23877v1 Announce Type: new Abstract: Logical reasoning serve as a central capability in LLMs and includes three main forms: deductive, inductive, and abductive reasoning. In this work, we s

KOMBO: Korean Character Representations Based on the Combination Rules of Subcharacters

ResearchDGX agent

arXiv:2604.23948v1 Announce Type: cross Abstract: The Korean writing system, extit{Hangeul}, has a unique character representation rigidly following the invention principles recorded in extit{Hunminje

LaDiR: Latent Diffusion Enhances LLMs for Text Reasoning

ResearchDGX agent

Large Language Models (LLMs) demonstrate their reasoning ability through chain-of-thought (CoT) generation. However, LLM’s autoregressive decoding may limit the ability to revisit and refine earlier t

Language Models Might Not Understand You: Evaluating Theory of Mind via Story Prompting

ResearchDGX agent

arXiv:2506.19089v5 Announce Type: replace-cross Abstract: We introduce StorySim, a programmable framework for synthetically generating stories to evaluate the theory of mind (ToM) and world modeling (

Large language model-enabled automated data extraction for concrete materials informatics

ResearchDGX agent

arXiv:2604.22938v1 Announce Type: cross Abstract: The promise of data-driven materials discovery remains constrained by the scarcity of large, high-quality, and accessible experimental datasets. Here,

LatentBurst: A Fast and Efficient Multi Frame Super-Resolution for Hexadeca-Bayer Pattern CIS images

ResearchDGX agent

arXiv:2604.23268v1 Announce Type: new Abstract: This paper introduces a novel multi frame super-resolution network (MFSR) for burst hexadeca Bayer pattern Contact Image Sensor (CIS) images, which incl

Layer Embedding Deep Fusion Graph Neural Network

ResearchDGX agent

arXiv:2604.23324v1 Announce Type: cross Abstract: Graph Neural Networks (GNNs) have demonstrated impressive performance in learning representations from graph-structured data. However, their message-p

Learning an Image Editing Model without Image Editing Pairs

ResearchDGX agent

arXiv:2510.14978v2 Announce Type: replace Abstract: Recent image editing models have achieved impressive results while following natural language editing instructions, but they rely on supervised fine

Learning-Augmented Robust Algorithmic Recourse

ResearchDGX agent

arXiv:2410.01580v3 Announce Type: replace Abstract: Algorithmic recourse provides individuals who receive undesirable outcomes from machine learning systems with minimum-cost improvements to achieve a

Learning Curves and Benign Overfitting of Spectral Algorithms in Large Dimensions

ResearchDGX agent

arXiv:2604.23212v1 Announce Type: cross Abstract: Existing large-dimensional theory for spectral algorithms resolves either the optimally tuned point or the interpolation limit, but leaves the under-r

Learning Evidence of Depression Symptoms via Prompt Induction

ResearchDGX agent

arXiv:2604.24376v1 Announce Type: new Abstract: Depression places substantial pressure on mental health services, and many people describe their experiences outside clinical settings in high-volume us

Learning Interpretable PDE Representations for Generative Reconstructions with Structured Sparsity

ResearchDGX agent

arXiv:2604.23867v1 Announce Type: new Abstract: Scientific measurements are often bottlenecked by suboptimal conditions, whether that be noise, incomplete spatial coverage, or limited resolution, rend

Learning Operators by Regularized Stochastic Gradient Descent with Operator-valued Kernels

ResearchDGX agent

arXiv:2504.18184v4 Announce Type: replace-cross Abstract: We consider a class of statistical inverse problems involving the estimation of a regression operator from a Polish space to a separable Hilbe

Learning to Orchestrate Agents in Natural Language with the Conductor Fugu Blog: https://sakana.ai/fugu-beta Paper: https://arxiv.org/abs/25…

ResearchDGX agent

Conductor is a method for orchestrating multiple AI agents through natural language instructions, enabling coordinated multi-agent systems where a central 'conductor' agent directs specialized agents

Learning Without Adversarial Training: A Physics-Informed Neural Network for Secure Power System State Estimation under False Data Injection Attacks

ResearchDGX agent

arXiv:2604.22784v1 Announce Type: new Abstract: State estimation is a cornerstone of power system control-center operations, and its robust operation is increasingly a cyber-physical security concern

Leveraging Human Feedback for Semantically-Relevant Skill Discovery

ResearchDGX agent

arXiv:2604.24127v1 Announce Type: cross Abstract: Unsupervised skill discovery in reinforcement learning aims to intrinsically motivate agents to discover diverse and useful behaviours. However, uncon

Leveraging Spatial Transcriptomics as Alternative to Manual Annotations for Deep Learning-Based Nuclei Analysis

ResearchDGX agent

arXiv:2604.23481v1 Announce Type: new Abstract: Deep learning-based nuclei segmentation and classification in pathology images typically rely on large-scale pixel-level manual annotations, which are c

Light 'em Up: Enabling Few-Shot Low-Light 3D Gaussian Splatting with Multi-Scale Explicit Retinex Illumination Decoupling

ResearchDGX agent

arXiv:2604.24053v1 Announce Type: new Abstract: Full 360^irc novel view synthesis under low-light conditions remains challenging. Insufficient illumination, noise amplification, and view-dependent pho

LILogic Net: Compact Logic Gate Networks with Learnable Connectivity for Efficient Hardware Deployment

ResearchDGX agent

arXiv:2511.12340v2 Announce Type: replace Abstract: Efficient machine learning deployment requires models that account for hardware constraints. Because binary logic gates are the fundamental primitiv

Live Knowledge Tracing: Real-Time Adaptation using Tabular Foundation Models

ResearchDGX agent

arXiv:2602.06542v3 Announce Type: replace Abstract: Deep knowledge tracing models have achieved significant breakthroughs in modeling student learning trajectories. However, these architectures requir

LLMind: Bio-inspired Training-free Adaptive Visual Representations for Vision-Language Models

ResearchDGX agent

arXiv:2603.14882v2 Announce Type: replace Abstract: Vision-Language Models (VLMs) typically assume a uniform spatial fidelity across the entire field of view of visual inputs, dedicating equal precisi

LoFi: Location-Aware Fine-Grained Representation Learning for Chest X-ray

ResearchDGX agent

arXiv:2603.19451v2 Announce Type: replace-cross Abstract: Fine-grained representation learning is crucial for retrieval and phrase grounding in chest X-rays, where clinically relevant findings are oft

Machine Learning for Network Attacks Classification and Statistical Evaluation of Adversarial Learning Methodologies for Synthetic Data Generation

ResearchDGX agent

arXiv:2603.17717v3 Announce Type: replace-cross Abstract: Supervised detection of network attacks has always been a critical part of network intrusion detection systems (NIDS). Nowadays, in a pivotal

Machine learning models for estimating counterfactuals in a single-arm inflammatory bowel disease study

ResearchDGX agent

arXiv:2604.23465v1 Announce Type: new Abstract: Single-arm trials accelerate study timelines by reducing the number of patients that must be recruited for a concurrent control group. However, these de

Making Dialogue Grounding Data Rich: A Three-Tier Data Synthesis Framework for Generalized Referring Expression Comprehension

ResearchDGX agent

arXiv:2512.02791v2 Announce Type: replace Abstract: Dialogue-Based Generalized Referring Expression Comprehension (GREC) requires models to ground the expression and unlimited targets in complex visua

Mammographic Lesion Segmentation with Lightweight Models: A Comparative Study

ResearchDGX agent

arXiv:2604.23899v1 Announce Type: new Abstract: Breast cancer is a leading cause of cancer-related mortality among women worldwide, with mammography as the primary screening tool. While deep learning

Maritime object classification with SAR imagery using quantum kernel methods

ResearchDGX agent

arXiv:2512.11367v2 Announce Type: replace-cross Abstract: Illegal, unreported, and unregulated (IUU) fishing causes global economic losses of 10-25 billion USD annually and undermines marine sustainab

MARRS: Masked Autoregressive Unit-based Reaction Synthesis

ResearchDGX agent

arXiv:2505.11334v4 Announce Type: replace Abstract: This work aims at a challenging task: human action-reaction synthesis, i.e., generating human reactions conditioned on the action sequence of anothe

Measuring Temporal Linguistic Emergence in Diffusion Language Models

ResearchDGX agent

arXiv:2604.23235v1 Announce Type: new Abstract: Diffusion language models expose an explicit denoising trajectory, making it possible to ask when different kinds of information become measurable durin

MedSpeak: A Knowledge Graph-Aided ASR Error Correction Framework for Spoken Medical QA

ResearchDGX agent

arXiv:2602.00981v2 Announce Type: replace-cross Abstract: Spoken question-answering (SQA) systems relying on automatic speech recognition (ASR) often struggle with accurately recognizing medical termi

MegaScale-Data: Scaling Dataloader for Multisource Large Foundation Model Training

ResearchDGX agent

arXiv:2504.09844v4 Announce Type: replace-cross Abstract: Modern frameworks for training large foundation models (LFMs) employ dataloaders in a data-parallel manner, with each loader processing a disj

MemeScouts@LT-EDI 2026: Asking the Right Questions -- Prompted Weak Supervision for Meme Hate Speech Detection

ResearchDGX agent

arXiv:2604.24179v1 Announce Type: cross Abstract: Detecting hate speech in memes is challenging due to their multimodal nature and subtle, culturally grounded cues such as sarcasm and context. While r

MeshLAM: Feed-Forward One-Shot Animatable Textured Mesh Avatar Reconstruction

ResearchDGX agent

arXiv:2604.22865v1 Announce Type: new Abstract: We introduce MeshLAM, a feed-forward framework for one-shot animatable mesh head reconstruction that generates high-fidelity, animatable 3D head avatars

Micro-Expression-Aware Avatar Fingerprinting via Inter-Frame Feature Differencing

ResearchDGX agent

arXiv:2604.23247v1 Announce Type: new Abstract: Avatar fingerprinting, i.e., verifying who drives a synthetic talking-head video rather than whether it is real, is a critical safeguard for authorized

MIMIC: A Generative Multimodal Foundation Model for Biomolecules

ResearchDGX agent

arXiv:2604.24506v1 Announce Type: new Abstract: Biological function emerges from coupled constraints across sequence, structure, regulation, evolution, and cellular context, yet most foundation models

MindTrellis: Co-Creating Knowledge Structures with AI through Interactive Visual Exploration

ResearchDGX agent

arXiv:2604.23129v1 Announce Type: cross Abstract: Knowledge workers face increasing challenges in synthesizing information from multiple documents into structured conceptual understanding. This proces

MirrorMark: A Distortion-Free Multi-Bit Watermark for Large Language Models

ResearchDGX agent

arXiv:2601.22246v2 Announce Type: replace-cross Abstract: As large language models (LLMs) become integral to applications such as question answering and content creation, reliable content attribution

mKG-RAG: Leveraging Multimodal Knowledge Graphs in Retrieval-Augmented Generation for Knowledge-intensive VQA

ResearchDGX agent

arXiv:2508.05318v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) has emerged as an effective paradigm for expanding the knowledge capacity of Multimodal Large Language Mo

Modeling Behavioral Intensity and Transitions for Generative Recommendation

ResearchDGX agent

arXiv:2604.24472v1 Announce Type: cross Abstract: Multi-behavior recommendation aims to predict user conversions by modeling various interaction types that carry distinct intent signals. Recently, gen

Modeling Induced Pleasure through Cognitive Appraisal Prediction via Multimodal Fusion

ResearchDGX agent

arXiv:2604.23753v1 Announce Type: new Abstract: Multimodal affective computing analyzes user-generated social media content to predict emotional states. However, a critical gap remains in understandin

Monocular Depth Estimation via Neural Network with Learnable Algebraic Group and Ring Structures

ResearchDGX agent

arXiv:2604.24328v1 Announce Type: new Abstract: Monocular depth estimation (MDE) has witnessed remarkable progress driven by Convolutional Neural Networks and transformer-based architectures. However,

Multi-Plane HyperX: A Low-Latency and Cost-Effective Network for Large-Scale AI and HPC Systems

ResearchDGX agent

arXiv:2604.23519v1 Announce Type: cross Abstract: Multi-plane architectures have become increasingly prevalent in the Fat-Tree networks of AI data centers. By leveraging multiple ports on a single net

Multi-Robot Motions in Milliseconds: Vector-Accelerated Primitives for Sampling-Based Planning

ResearchDGX agent

arXiv:2604.23960v1 Announce Type: new Abstract: In this paper, we extend the recent Vector-Accelerated Motion Planning (VAMP) framework to multi-robot motion planning (MRMP). We develop two vector-acc

Multi-Scale Contrastive Learning for Video Temporal Grounding

ResearchDGX agent

arXiv:2412.07157v3 Announce Type: replace Abstract: Temporal grounding, which localizes video moments related to a natural language query, is a core problem of vision-language learning and video under

Multi-scale Dynamic Wake Modeling of Floating Offshore Wind Turbines via Fourier Neural Operators and Physics-Informed Neural Networks

ResearchDGX agent

arXiv:2604.23937v1 Announce Type: cross Abstract: Multi-scale dynamic wake prediction is essential for the real-time control and performance optimization of floating offshore wind turbines (FOWTs). In

Multi-view Graph Convolutional Network with Fully Leveraging Consistency via Granular-ball-based Topology Construction, Feature Enhancement and Interactive Fusion

ResearchDGX agent

arXiv:2603.26729v2 Announce Type: replace-cross Abstract: The effective utilization of consistency is crucial for multi-view learning. GCNs leverage node connections to propagate information across th

Multimodal QUD: Inquisitive Questions from Scientific Figures

ResearchDGX agent

arXiv:2604.23733v1 Announce Type: new Abstract: Asking inquisitive questions while reading, and looking for their answers, is an important part in human discourse comprehension, curiosity, and creativ

← Previous
1…260261262263264…317
Next →